← 返回
内容创作 Key 中文

Alicloud Ai Entry Modelstudio

Route Alibaba Cloud Model Studio requests to the right local skill (Qwen Image, Qwen Image Edit, Wan Video, Wan R2V, Qwen TTS, Qwen ASR and advanced TTS vari...
将阿里云 Model Studio 请求路由至相应本地技能(千问图像、千问图像编辑、万视频、万 R2V、千问 TTS、千问 ASR 及高级 TTS 变体...)
cinience
内容创作 clawhub v1.0.3 2 版本 99935.4 Key: 需要
★ 0
Stars
📥 1,548
下载
💾 109
安装
2
版本
#latest

概述

Category: task

Alibaba Cloud Model Studio Entry (Routing)

Route requests to existing local skills to avoid duplicating model/parameter details.

Prerequisites

  • Install SDK (virtual environment recommended to avoid PEP 668 restrictions):
python3 -m venv .venv
. .venv/bin/activate
python -m pip install dashscope
  • Configure DASHSCOPE_API_KEY (environment variable preferred; or dashscope_api_key in ~/.alibabacloud/credentials).

Routing Table (currently supported in this repo)

NeedTarget skill
------
Text-to-image / image generationskills/ai/image/alicloud-ai-image-qwen-image/
Image editingskills/ai/image/alicloud-ai-image-qwen-image-edit/
Text-to-video / image-to-video (i2v)skills/ai/video/alicloud-ai-video-wan-video/
Reference-to-video (r2v)skills/ai/video/alicloud-ai-video-wan-r2v/
Text-to-speech (TTS)skills/ai/audio/alicloud-ai-audio-tts/
Speech recognition/transcription (ASR)skills/ai/audio/alicloud-ai-audio-asr/
Realtime speech recognitionskills/ai/audio/alicloud-ai-audio-asr-realtime/
Realtime TTSskills/ai/audio/alicloud-ai-audio-tts-realtime/
Live speech translationskills/ai/audio/alicloud-ai-audio-livetranslate/
CosyVoice voice cloneskills/ai/audio/alicloud-ai-audio-cosyvoice-voice-clone/
CosyVoice voice designskills/ai/audio/alicloud-ai-audio-cosyvoice-voice-design/
Voice cloneskills/ai/audio/alicloud-ai-audio-tts-voice-clone/
Voice designskills/ai/audio/alicloud-ai-audio-tts-voice-design/
Omni multimodal interactionskills/ai/multimodal/alicloud-ai-multimodal-qwen-omni/
Visual reasoningskills/ai/multimodal/alicloud-ai-multimodal-qvq/
Text embeddingsskills/ai/search/alicloud-ai-search-text-embedding/
Rerankskills/ai/search/alicloud-ai-search-rerank/
Vector retrievalskills/ai/search/alicloud-ai-search-dashvector/ or skills/ai/search/alicloud-ai-search-opensearch/ or skills/ai/search/alicloud-ai-search-milvus/
Document understandingskills/ai/text/alicloud-ai-text-document-mind/
Video editingskills/ai/video/alicloud-ai-video-wan-edit/
Model list crawl/updateskills/ai/misc/alicloud-ai-misc-crawl-and-skill/

When Not Matched

  • Clarify model capability and input/output type first.
  • If capability is missing in repo, add a new skill first.

Common Missing Capabilities In This Repo (remaining gaps)

  • text generation/chat (LLM)
  • multimodal embeddings
  • OCR-specialized extraction and image translation
  • virtual try-on / digital human / advanced video personas
  • For multimodal/ASR download failures, prefer public URLs listed above.
  • For ASR parameter errors, use data URI in input_audio.data.
  • For multimodal embedding 400, ensure input.contents is an array.

Async Task Polling Template (video/long-running tasks)

When X-DashScope-Async: enable returns task_id, poll as follows:

GET https://dashscope.aliyuncs.com/api/v1/tasks/<task_id>
Authorization: Bearer $DASHSCOPE_API_KEY

Example result fields (success):

{
  "output": {
    "task_status": "SUCCEEDED",
    "video_url": "https://..."
  }
}

Notes:

  • Recommended polling interval: 15-20 seconds, max 10 attempts.
  • After success, download output.video_url.

Clarifying questions (ask when uncertain)

  1. Are you working with text, image, audio, or video?
  2. Is this generation, editing/understanding, or retrieval?
  3. Do you need speech (TTS/ASR/live translate) or retrieval (embedding/rerank/vector DB)?
  4. Do you want runnable SDK scripts or just API/parameter guidance?

References

  • Model list and links:output/alicloud-model-studio-models-summary.md
  • API/parameters/examples: see target sub-skill SKILL.md and references/*.md
  • Official source list:references/sources.md

Validation

mkdir -p output/alicloud-ai-entry-modelstudio
echo "validation_placeholder" > output/alicloud-ai-entry-modelstudio/validate.txt

Pass criteria: command exits 0 and output/alicloud-ai-entry-modelstudio/validate.txt is generated.

Output And Evidence

  • Save artifacts, command outputs, and API response summaries under output/alicloud-ai-entry-modelstudio/.
  • Include key parameters (region/resource id/time range) in evidence files for reproducibility.

Workflow

1) Confirm user intent, region, identifiers, and whether the operation is read-only or mutating.

2) Run one minimal read-only query first to verify connectivity and permissions.

3) Execute the target operation with explicit parameters and bounded scope.

4) Verify results and save output/evidence files.

版本历史

共 2 个版本

  • v1.0.3 当前
    2026-03-28 22:48 安全 安全
  • v1.0.2
    2026-03-11 11:09

安全检测

腾讯云安全 (Keen)

安全,无风险
查看报告

腾讯云安全 (Sanbu)

安全,无风险
查看报告

🔗 相关推荐

content-creation

AdMapix

fly0pants
广告情报与应用数据分析助手,支持搜索广告素材、分析应用排名、下载量、收入及市场洞察,用于广告素材和竞品分析。
★ 295 📥 136,547
ai-intelligence

Volcengine Ai Audio Tts

cinience
在火山引擎音频服务上进行文本转语音生成。适用于需要配音、多语言语音输出、声音选择或TTS故障排除的场景。
★ 1 📥 2,186
content-creation

Humanizer

biostartechnology
消除AI写作痕迹,使文本更自然真实。基于维基百科"AI写作特征"指南,识别并修正夸张象征、宣传用语、肤浅-ing分析、模糊归因、破折号滥用、三项排比、AI词汇、负面平行结构及冗长连接词等模式。
★ 861 📥 200,188