ModelInfo
English
All models

GLM-TTS-Clone

Voice cloning model that learns a speaker's timbre from a 3-second audio sample.

Zhipu AIglm-tts-clone2025-12-11Report incorrect data

Specs

Context
—
Max output
—
Input price
—
Output price
¥6/ request
Released
2025-12-11

Context

Not published

Pricing · CNY / request

Input
—
Output
¥6
Cache read
—
Cache write
—

Modalities

Input
AudioText
Output
Audio

API

API types
audio

Reasoning

Not published

Info

Status
Active
Released
2025-12-11
Knowledge cutoff
—

Features

Not published
Source · official docsVerified 2026-10-11

How to call

1 provider

zhipumodel = glm-tts-clone

audio
POSThttps://open.bigmodel.cn/api/paas/v4/voice/clone

JSON

Standard format
glm-tts-clone.json
{
  "id": "glm-tts-clone",
  "object": "model",
  "created": 1765411200,
  "owned_by": "zhipu",
  "name": "GLM-TTS-Clone",
  "api": {
    "types": [
      "audio"
    ]
  },
  "limits": {
    "context": null,
    "input": null,
    "output": null
  },
  "modalities": {
    "input": [
      "audio",
      "text"
    ],
    "output": [
      "audio"
    ]
  },
  "reasoning": {
    "supported": null,
    "efforts": []
  },
  "pricing": {
    "currency": "CNY",
    "unit": "request",
    "input": null,
    "output": 6,
    "cache_read": null,
    "cache_write": null
  },
  "features": [],
  "info": {
    "status": "active",
    "release_date": "2025-12-11",
    "knowledge_cutoff": null,
    "description": "Voice cloning model that learns a speaker's timbre from a 3-second audio sample.",
    "docs": "https://docs.bigmodel.cn/cn/guide/models/sound-and-video/glm-tts-clone",
    "verified_at": "2026-10-11"
  }
}

Official sources

5