logo
RunComfy
  • ComfyUI
  • 训练器新
  • 模型
  • API
  • 定价
discord logo
模型
探索
所有模型
资源库
生成记录
模型 API
API 文档
API 密钥
账户
使用情况

Wan 3.0 Reference To Video:Models and API 上的多模态参考视频生成 | RunComfy

wan-ai/wan-3.0/reference-to-video

Wan 3.0 Reference To Video 将提示与图像、视频和音频参考相结合,以构建具有一致主题、动作和声音的连贯剪辑。

目录

1. 快速开始2. 身份验证3. API 参考提交请求查询请求状态获取请求结果取消请求4. 文件输入托管文件(URL)5. 数据结构输入结构输出结构

1. 快速开始

使用 RunComfy 的 API 运行 wan-ai/wan-3.0/reference-to-video。 可接受的输入与输出请参阅模型的 数据结构说明。

curl --request POST \
  --url https://model-api.runcomfy.net/v1/models/wan-ai/wan-3.0/reference-to-video \
  --header "Content-Type: application/json" \
  --header "Authorization: Bearer <token>" \
  --data '{
    "prompt": "Image 1 walks slowly through a bright modern art gallery, then turns toward the camera and smiles; soft natural daylight, elegant handheld motion, cinematic, subtle ambient sound."
  }'

2. 身份验证

将 YOUR_API_TOKEN 环境变量设置为您的 API 密钥(在 个人资料中管理密钥),并在每个请求的 Authorization 标头中以 Bearer 令牌形式携带: Authorization: Bearer $YOUR_API_TOKEN。

3. API 参考

提交请求

提交异步生成任务,将立即获得 request_id 以及用于查询状态、获取结果和取消的 URL。

curl --request POST \
  --url https://model-api.runcomfy.net/v1/models/wan-ai/wan-3.0/reference-to-video \
  --header "Content-Type: application/json" \
  --header "Authorization: Bearer <token>" \
  --data '{
    "prompt": "Image 1 walks slowly through a bright modern art gallery, then turns toward the camera and smiles; soft natural daylight, elegant handheld motion, cinematic, subtle ambient sound."
  }'

查询请求状态

根据 request_id 获取当前状态("in_queue"、"in_progress"、"completed" 或 "cancelled")。

curl --request GET \
  --url https://model-api.runcomfy.net/v1/requests/{request_id}/status \
  --header "Authorization: Bearer <token>"

获取请求结果

获取指定 request_id 的最终输出与元数据;若任务未完成,响应会返回当前状态,便于继续轮询。

curl --request GET \
  --url https://model-api.runcomfy.net/v1/requests/{request_id}/result \
  --header "Authorization: Bearer <token>"

取消请求

通过 request_id 取消排队中的任务;进行中的任务无法取消。

curl --request POST \
  --url https://model-api.runcomfy.net/v1/requests/{request_id}/cancel \
  --header "Authorization: Bearer <token>"

4. 文件输入

托管文件(URL)

请提供可公网访问的 HTTPS 地址。确保目标主机允许服务端拉取(无需登录或 Cookie)、未被限流或拦截爬虫。建议:图片 ≤ 50 MB(约 4K),视频 ≤ 100 MB(约 720p 下 2–5 分钟)。私有资源请使用稳定或预签名 URL。

5. 数据结构

输入结构

{
  "type": "object",
  "title": "输入结构",
  "required": [
    "prompt"
  ],
  "properties": {
    "prompt": {
      "title": "提示词",
      "description": "描述场景、主题、动作、摄像机移动、灯光和风格。按顺序参考图像 1、视频 1、音频 1。",
      "type": "string",
      "default": "Image 1 walks slowly through a bright modern art gallery, then turns toward the camera and smiles; soft natural daylight, elegant handheld motion, cinematic, subtle ambient sound."
    },
    "reference_images": {
      "title": "参考图片",
      "description": "最多 10 个参考图像,以确保主题、物体或场景的一致性。至少需要一份参考资料(图像、视频或音频)。",
      "type": "array",
      "items": {
        "type": "string",
        "format": "image_uri"
      },
      "default": [
        "https://playgrounds-storage-public.runcomfy.net/tools/7415/media-files/input-promo-ref.webp"
      ]
    },
    "reference_videos": {
      "title": "参考视频",
      "description": "最多 5 个参考视频(MP4 或 MOV,每个 1-15 秒,总计不超过 15 秒)用于动作或场景指导。",
      "type": "array",
      "items": {
        "type": "string",
        "format": "video_uri"
      },
      "default": []
    },
    "reference_audios": {
      "title": "参考音频",
      "description": "最多 5 个参考音频剪辑(总计不超过 15 秒)来指导声音或计时。",
      "type": "array",
      "items": {
        "type": "string",
        "format": "audio_uri"
      },
      "default": []
    },
    "resolution": {
      "title": "分辨率",
      "description": "输出分辨率层。使用 480p 进行快速草稿,使用 1080p 获得更高质量的输出。",
      "type": "string",
      "enum": [
        "480p",
        "720p",
        "1080p"
      ],
      "default": "720p"
    },
    "aspect_ratio": {
      "title": "宽高比",
      "description": "生成视频的输出宽高比。",
      "type": "string",
      "enum": [
        "16:9",
        "9:16",
        "1:1",
        "4:3",
        "3:4",
        "adaptive"
      ],
      "default": "16:9"
    },
    "duration": {
      "title": "持续时间(秒)",
      "description": "生成视频的长度(以秒为单位)。范围为 2-30。开始时进行短距离迭代,然后在运动看起来正确后增加。",
      "type": "integer",
      "minimum": 2,
      "maximum": 30,
      "default": 5
    },
    "thinking_mode": {
      "title": "深度思考模式",
      "description": "为具有多种动作或构图要求的复杂场景提供更深入的提示解释。",
      "type": "boolean",
      "default": false
    },
    "enable_audio": {
      "title": "生成音频",
      "description": "打开时,输出视频包括同步音轨。关闭以获得无声剪辑。",
      "type": "boolean",
      "default": true
    },
    "seed": {
      "title": "种子",
      "description": "随机种子以获得可重复的结果。范围为 0 到 2147483647。",
      "type": "integer"
    }
  }
}

输出结构

{
  "output": {
    "type": "object",
    "properties": {
      "image": {
        "type": "string",
        "format": "uri",
        "description": "单张图片 URL"
      },
      "video": {
        "type": "string",
        "format": "uri",
        "description": "单个视频 URL"
      },
      "images": {
        "type": "array",
        "description": "多张图片 URL",
        "items": {
          "type": "string",
          "format": "uri"
        }
      },
      "videos": {
        "type": "array",
        "description": "多个视频 URL",
        "items": {
          "type": "string",
          "format": "uri"
        }
      }
    }
  }
}
关注我们
  • 领英
  • Facebook
  • Instagram
  • Twitter
支持
  • Discord
  • 电子邮件
  • 系统状态
  • 附属
视频模型
  • MiniMax H3 Open
  • FLUX 3 Image to Video
  • MiniMax H3 Open Image to Video
  • Wan 2.6 Flash
  • Happy Horse 1.1 reference to video
  • Seedance 1.5 Pro Text to Video
  • 查看所有模型 →
图像模型
  • Seedream 5.0 Pro Image Edit
  • Flux 2 Flash Edit
  • Nano Banana Pro
  • seedream 4.0
  • GPT Image 2
  • Qwen Image Edit 2511 LoRA
  • 查看所有模型 →
法律
  • 服务条款
  • 隐私政策
  • Cookie 政策
RunComfy
版权 2026 RunComfy. 保留所有权利。

RunComfy 是首选的 ComfyUI 平台,提供 ComfyUI 在线 环境和服务,以及 ComfyUI 工作流 具有惊艳的视觉效果。 RunComfy还提供 AI Models, 帮助艺术家利用最新的AI工具创作出令人惊叹的艺术作品。