logo
RunComfy
  • ComfyUI
  • 训练器新
  • 模型
  • API
  • 定价
discord logo
模型
探索
所有模型
资源库
生成记录
模型 API
API 文档
API 密钥
账户
使用情况

Seed Audio 1.0:支持预设声音与参考音频的文生音频模型和 API | RunComfy

bytedance/seed-audio-1.0/text-to-audio

在 RunComfy 上使用 Seed Audio 1.0,根据文字、参考音频片段或图像生成自然的语音音频。支持预设声音,并可调整语速、音高和输出格式,可在浏览器中使用或通过 API 调用。

目录

1. 快速开始2. 身份验证3. API 参考提交请求查询请求状态获取请求结果取消请求4. 文件输入托管文件(URL)5. 数据结构输入结构输出结构

1. 快速开始

使用 RunComfy 的 API 运行 bytedance/seed-audio-1.0/text-to-audio。 可接受的输入与输出请参阅模型的 数据结构说明。

curl --request POST \
  --url https://model-api.runcomfy.net/v1/models/bytedance/seed-audio-1.0/text-to-audio \
  --header "Content-Type: application/json" \
  --header "Authorization: Bearer <token>" \
  --data '{
    "prompt": "Welcome back to the late-night show. Settle in, pour something warm, and let's ease into the next track together."
  }'

2. 身份验证

将 YOUR_API_TOKEN 环境变量设置为您的 API 密钥(在 个人资料中管理密钥),并在每个请求的 Authorization 标头中以 Bearer 令牌形式携带: Authorization: Bearer $YOUR_API_TOKEN。

3. API 参考

提交请求

提交异步生成任务,将立即获得 request_id 以及用于查询状态、获取结果和取消的 URL。

curl --request POST \
  --url https://model-api.runcomfy.net/v1/models/bytedance/seed-audio-1.0/text-to-audio \
  --header "Content-Type: application/json" \
  --header "Authorization: Bearer <token>" \
  --data '{
    "prompt": "Welcome back to the late-night show. Settle in, pour something warm, and let's ease into the next track together."
  }'

查询请求状态

根据 request_id 获取当前状态("in_queue"、"in_progress"、"completed" 或 "cancelled")。

curl --request GET \
  --url https://model-api.runcomfy.net/v1/requests/{request_id}/status \
  --header "Authorization: Bearer <token>"

获取请求结果

获取指定 request_id 的最终输出与元数据;若任务未完成,响应会返回当前状态,便于继续轮询。

curl --request GET \
  --url https://model-api.runcomfy.net/v1/requests/{request_id}/result \
  --header "Authorization: Bearer <token>"

取消请求

通过 request_id 取消排队中的任务;进行中的任务无法取消。

curl --request POST \
  --url https://model-api.runcomfy.net/v1/requests/{request_id}/cancel \
  --header "Authorization: Bearer <token>"

4. 文件输入

托管文件(URL)

请提供可公网访问的 HTTPS 地址。确保目标主机允许服务端拉取(无需登录或 Cookie)、未被限流或拦截爬虫。建议:图片 ≤ 50 MB(约 4K),视频 ≤ 100 MB(约 720p 下 2–5 分钟)。私有资源请使用稳定或预签名 URL。

5. 数据结构

输入结构

{
  "type": "object",
  "title": "输入结构",
  "required": [
    "prompt"
  ],
  "properties": {
    "prompt": {
      "title": "提示词",
      "description": "要合成的文字。请按顺序使用 @Audio1、@Audio2、@Audio3 引用参考音频片段。",
      "type": "string",
      "default": "Welcome back to the late-night show. Settle in, pour something warm, and let's ease into the next track together."
    },
    "voice": {
      "title": "声音",
      "description": "用于合成的预设声音。",
      "type": "string",
      "enum": [
        "vivi_mixed_en_zh_ja_es_id",
        "mindy_en_es_id_pt_zh",
        "kian_en_zh",
        "cedric_en_zh",
        "sophie_en_zh",
        "jean_en_zh",
        "magnus_en_zh",
        "mabel_en_zh",
        "nadia_en_zh",
        "opal_en_zh",
        "pearl_en_zh",
        "quentin_en_zh",
        "corinne_mixed_en_zh",
        "esther_mixed_en_zh",
        "lyla_mixed_en_zh",
        "tracy_es_zh",
        "sandy_es_mixed_en_zh",
        "felix_zh",
        "celeste_zh",
        "monkey_king_zh"
      ],
      "default": "vivi_mixed_en_zh_ja_es_id"
    },
    "audio_urls": {
      "title": "参考音频 URL",
      "description": "最多可添加 3 段参考音频,并在提示词中以 @Audio1、@Audio2、@Audio3 引用。每段最长 30s、最大 10MB,支持 wav/mp3/pcm/ogg_opus。",
      "type": "array",
      "items": {
        "type": "string",
        "format": "audio_uri"
      },
      "maxItems": 3
    },
    "image_url": {
      "title": "参考图像 URL",
      "description": "一张参考图像(jpeg/png/webp,最大 10MB)。不能与参考音频同时使用。",
      "type": "string"
    },
    "output_format": {
      "title": "输出格式",
      "description": "输出音频的容器格式。",
      "type": "string",
      "enum": [
        "wav",
        "mp3",
        "pcm",
        "ogg_opus"
      ],
      "default": "mp3"
    },
    "sample_rate": {
      "title": "采样率 (Hz)",
      "description": "输出音频的采样率,以 Hz 为单位。",
      "type": "integer",
      "enum": [
        8000,
        16000,
        24000,
        32000,
        44100,
        48000
      ],
      "default": 24000
    },
    "speed": {
      "title": "语速",
      "description": "朗读速度。1.0 为正常速度,0.5 为半速,2.0 为双倍速度。",
      "type": "number",
      "minimum": 0.5,
      "maximum": 2,
      "default": 1
    },
    "volume": {
      "title": "音量",
      "description": "输出响度。1.0 为正常音量,0.5 为一半,2.0 为两倍。",
      "type": "number",
      "minimum": 0.5,
      "maximum": 2,
      "default": 1
    },
    "pitch": {
      "title": "音高",
      "description": "以半音为单位调整声音音高。0 为正常音高,-12 降低一个八度,12 升高一个八度。",
      "type": "integer",
      "minimum": -12,
      "maximum": 12,
      "default": 0
    }
  }
}

输出结构

{
  "output": {
    "type": "object",
    "properties": {
      "image": {
        "type": "string",
        "format": "uri",
        "description": "单张图片 URL"
      },
      "video": {
        "type": "string",
        "format": "uri",
        "description": "单个视频 URL"
      },
      "images": {
        "type": "array",
        "description": "多张图片 URL",
        "items": {
          "type": "string",
          "format": "uri"
        }
      },
      "videos": {
        "type": "array",
        "description": "多个视频 URL",
        "items": {
          "type": "string",
          "format": "uri"
        }
      }
    }
  }
}
关注我们
  • 领英
  • Facebook
  • Instagram
  • Twitter
支持
  • Discord
  • 电子邮件
  • 系统状态
  • 附属
视频模型
  • MiniMax H3 Open
  • FLUX 3 Image to Video
  • MiniMax H3 Open Image to Video
  • Wan 2.6 Flash
  • Happy Horse 1.1 reference to video
  • Seedance 1.5 Pro Text to Video
  • 查看所有模型 →
图像模型
  • Seedream 5.0 Pro Image Edit
  • Flux 2 Flash Edit
  • Nano Banana Pro
  • seedream 4.0
  • GPT Image 2
  • Qwen Image Edit 2511 LoRA
  • 查看所有模型 →
法律
  • 服务条款
  • 隐私政策
  • Cookie 政策
RunComfy
版权 2026 RunComfy. 保留所有权利。

RunComfy 是首选的 ComfyUI 平台,提供 ComfyUI 在线 环境和服务,以及 ComfyUI 工作流 具有惊艳的视觉效果。 RunComfy还提供 AI Models, 帮助艺术家利用最新的AI工具创作出令人惊叹的艺术作品。