本文へ移動
cccskills
無料GitHub で公開

byted-vod-process-tools

Volcengine VOD audio and video processing tools skill. Use when users need VOD-based media processing or editing: upload local/URL media, stitch videos, trim clips, flip frames, change playback speed, create image-to-video, compose audio/video, extract audio, mix audio, separate vocals/accompaniment, denoise audio, enhance quality, AI super-resolution, frame interpolation, ASR speech-to-text, OCR text extraction, subtitle removal, subtitle embedding, scene slicing, portrait/green-screen matting, highlight extraction, comic style transfer, video translation, drama recap narration, drama script restoration, media info lookup, or playback URL retrieval. The skill submits async VOD jobs, polls task status, and returns generated output links. Not for pure text generation, real-time streaming, or source-free generative video creation.

インストール方法を見る

含まれるファイル(68)

  • SKILL.md11.8 KB
  • LICENSE9.9 KB
  • references/00-billing-instructions.md1.5 KB
  • references/01-stitching.md1.8 KB
  • references/02-clipping.md995 B
  • references/03-flip.md932 B
  • references/04-speedup.md1.1 KB
  • references/05-image-to-video.md1.7 KB
  • references/06-compile.md1.6 KB
  • references/07-extract-audio.md686 B
  • references/08-mix-audios.md943 B
  • references/09-add-sub-video.md1.9 KB
  • references/10-voice-separation.md1.3 KB
  • references/11-noise-reduction.md1.1 KB
  • references/12-quality-enhance.md1004 B
  • references/13-super-resolution.md1.5 KB
  • references/14-interlacing.md1023 B
  • references/15-asr-speech-to-text.md1.7 KB
  • references/16-ocr-text-extract.md845 B
  • references/17-subtitle-removal.md888 B
  • references/18-add-subtitle.md2.9 KB
  • references/19-intelligent-slicing.md1.2 KB
  • references/20-portrait-matting.md1.0 KB
  • references/21-green-screen.md1010 B
  • references/22-comic-style.md2.5 KB
  • references/23-highlight.md3.6 KB
  • references/24-video-translation.md9.6 KB
  • references/25-drama-recap.md7.4 KB
  • references/26-drama-script.md4.3 KB
  • references/27-get-media-info.md1.5 KB
  • scripts/add_subtitle.py2.1 KB
  • scripts/api_manage.py53.3 KB
  • scripts/asr_speech_to_text.py1.7 KB
  • scripts/clipping.py1.6 KB
  • scripts/comic_style.py5.9 KB
  • scripts/compile.py2.0 KB
  • scripts/drama_recap.py11.8 KB
  • scripts/drama_script.py7.0 KB
  • scripts/extract_audio.py1.5 KB
  • scripts/flip.py1.5 KB
  • scripts/get_media_info.py4.4 KB
  • scripts/green_screen.py1.8 KB
  • scripts/highlight.py5.3 KB
  • scripts/image_to_video.py1.8 KB
  • scripts/intelligent_slicing.py1.8 KB
  • scripts/interlacing.py2.0 KB
  • scripts/list_translation.py4.7 KB
  • scripts/log_utils.py1.5 KB
  • scripts/mix_audios.py1.5 KB
  • scripts/noise_reduction.py1.6 KB
  • scripts/ocr_text_extract.py1.5 KB
  • scripts/poll_media.py1.5 KB
  • scripts/poll_translation.py3.3 KB
  • scripts/poll_vcreative.py1.3 KB
  • scripts/portrait_matting.py1.9 KB
  • scripts/quality_enhance.py1.8 KB
  • scripts/speedup.py2.0 KB
  • scripts/stitching.py2.1 KB
  • scripts/subtitle_removal.py1.7 KB
  • scripts/super_resolution.py2.5 KB
  • scripts/upload_media.py7.3 KB
  • scripts/video_translation.py14.8 KB
  • scripts/vod_api_constants.py2.1 KB
  • scripts/vod_common.py5.0 KB
  • scripts/vod_local_upload.py19.6 KB
  • scripts/vod_transport.py4.9 KB
  • scripts/voice_separation.py1.6 KB
  • scripts/volc_request.py4.8 KB

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

Volcengine VOD Tools


前置条件

  • Python:确认 python --version ≥ 3.6
  • 环境变量(必需,也可通过工作目录下的 .env 文件配置,脚本会自动加载):
    • VOLCENGINE_ACCESS_KEY — 火山引擎 Access Key
    • VOLCENGINE_SECRET_KEY — 火山引擎 Secret Key
    • VOD_SPACE_NAME — VOD 空间名称
  • 依赖:脚本依赖 python-dotenv

参数传入方式

所有脚本支持两种 JSON 参数传入方式:

  1. 内联 JSON(适合简单参数):python script.py '{"key":"value"}'
  2. 文件引用(推荐,避免 shell 转义问题):python script.py @params.json

@ 前缀表示从文件读取 JSON 内容,文件路径相对于当前工作目录。


结果交付规则

  • 提交异步任务成功后会返回异步任务id,字段为 VCCreativeId 或 TaskId,在给用户交付最终产物时,必须包含异步任务id
  • 在展示最终产物链接时,禁止随意修改链接内容
  • 优先将产物链接提供给用户

工作流程

1) 识别输入视频类型(必要时先上传拿 vid://...)

后续所有处理脚本优先使用 VOD 侧资源引用:

  • Vid:vid://vxxxx(或部分脚本接受裸 vxxxx 并自动补 vid://)
  • DirectUrl / FileName:directurl://<vod_file_name>(媒体类任务用 DirectUrl 时会要求 FileName + SpaceName)

当用户提供的是以下输入之一,需要先执行上传逻辑,拿到 Vid 后再继续:

  • 本地文件路径:如 /path/to/a.mp4
  • http/https 链接:如 https://example.com/a.mp4(会走 URL 拉取上传,并轮询上传结果)

统一用 scripts/upload_media.py:

python <SKILL_DIR>/scripts/upload_media.py "<local_file_path_or_http_url>" [space_name]

脚本输出中 Source 字段即 vid://...,可直接作为后续处理输入。

安全限制:本地文件上传仅允许 workspace/、userdata/ 和 /tmp 目录下的文件。

2) 识别用户意图 → 选择对应处理脚本

根据用户需求,按以下决策树选择脚本:

用户意图脚本
多个视频/音频合成一个(顺序拼接)stitching
截取视频/音频的某个时间片段clipping
加速/慢放/变速speedup
镜像/上下翻转/左右翻转flip
多张图片串联生成视频image_to_video
替换/叠加视频的背景音乐compile
只要视频里的音频轨extract_audio
多条音频同时叠加播放(混音)mix_audios
分离人声和伴奏/背景音voice_separation
去除环境噪音/电流杂音/风噪noise_reduction
模糊/低画质视频修复(压缩伪影/噪点/划痕)quality_enhance
低分辨率视频提升(如 720P→1080P)super_resolution
低帧率视频插帧提升流畅度(如 30fps→60fps)interlacing
语音识别/ASR/提取视频中的文字对白asr_speech_to_text
OCR 文字提取/识别视频中的屏幕文字ocr_text_extract
擦除视频硬字幕subtitle_removal
给视频添加/嵌入字幕(烧录字幕)add_subtitle
视频场景分割/智能切片intelligent_slicing
人像抠图/人像分割portrait_matting
绿幕抠像/绿屏抠像green_screen
AI 漫剧转绘(漫画风/3D卡通风格)comic_style
短剧高光剪辑/精彩片段提取highlight
AI 视频翻译(字幕/语音/面容翻译)video_translation
查询翻译项目状态/重启翻译轮询poll_translation
查询翻译项目列表list_translation
AI 解说视频生成(短剧解说/二创)drama_recap
AI 剧本还原(视频转结构化剧本)drama_script
查询媒资信息(Vid 详情+播放地址)get_media_info

3) 构造参数并执行

视频编辑类

脚本用途详细参数
stitching.py '<json>'视频/音频拼接references/01-stitching.md
clipping.py '<json>'视频/音频裁剪references/02-clipping.md
flip.py '<json>'视频翻转references/03-flip.md
speedup.py video '<json>'视频变速references/04-speedup.md
speedup.py audio '<json>'音频变速references/04-speedup.md
image_to_video.py '<json>'图片转视频references/05-image-to-video.md
compile.py '<json>'音视频合成references/06-compile.md
extract_audio.py '<json>'提取音轨references/07-extract-audio.md
mix_audios.py '<json>'混音references/08-mix-audios.md

媒体处理类

脚本用途详细参数
voice_separation.py '<json>'人声分离references/10-voice-separation.md
noise_reduction.py '<json>'音频降噪references/11-noise-reduction.md
quality_enhance.py '<json>'综合画质修复references/12-quality-enhance.md
super_resolution.py '<json>'AI 超分辨率references/13-super-resolution.md
interlacing.py '<json>'智能补帧references/14-interlacing.md

AI 内容分析类

脚本用途详细参数
asr_speech_to_text.py '<json>'语音识别 ASRreferences/15-asr-speech-to-text.md
ocr_text_extract.py '<json>'OCR 文字提取references/16-ocr-text-extract.md
subtitle_removal.py '<json>'硬字幕擦除references/17-subtitle-removal.md
add_subtitle.py '<json>'添加嵌入字幕references/18-add-subtitle.md
intelligent_slicing.py '<json>'智能场景分割references/19-intelligent-slicing.md
portrait_matting.py '<json>'人像抠图references/20-portrait-matting.md
green_screen.py '<json>'绿幕抠像references/21-green-screen.md
highlight.py '<json>'短剧高光剪辑references/23-highlight.md
get_media_info.py '<json>'媒资信息查询references/27-get-media-info.md

AI 内容生成类

脚本用途详细参数
comic_style.py '<json>'AI 漫剧转绘references/22-comic-style.md
video_translation.py '<json>'AI 视频翻译references/24-video-translation.md
drama_recap.py '<json>'AI 解说视频生成references/25-drama-recap.md
drama_script.py '<json>'AI 剧本还原references/26-drama-script.md

重启轮询

脚本用途
poll_vcreative.py <task_id>重启编辑类任务轮询
poll_media.py <task_type> <RunId>重启媒体处理类任务轮询
poll_translation.py <ProjectId>重启翻译任务轮询

超时响应中的 resume_hint.command 字段包含可直接复制执行的重启命令。


示例

# 本地文件先上传拿到 vid(后续脚本统一用 vid://... 作为输入)
python <SKILL_DIR>/scripts/upload_media.py "/path/to/local.mp4" my_space

# 拼接两个视频,加转场
python <SKILL_DIR>/scripts/stitching.py \
  '{"type":"video","videos":["vid://v0001","vid://v0002"],"transitions":["1182359"]}'

# 使用 @file.json 传参(推荐,避免转义问题)
python <SKILL_DIR>/scripts/stitching.py @params.json

# 人声分离(注意 type 首字母大写)
python <SKILL_DIR>/scripts/voice_separation.py '{"type":"Vid","video":"v0310abc"}'

# 超分到 1080P
python <SKILL_DIR>/scripts/super_resolution.py '{"type":"Vid","video":"v0310xyz","Res":"1080p"}'

# ASR 语音识别
python <SKILL_DIR>/scripts/asr_speech_to_text.py '{"type":"Vid","video":"v0310abc"}'

# 短剧高光剪辑
python <SKILL_DIR>/scripts/highlight.py '{"Vids":["v023xxx","v024xxx"]}'

# AI 视频翻译(中文→英文)
python <SKILL_DIR>/scripts/video_translation.py '{"Vid":"v0d225gxxx","SourceLanguage":"zh","TargetLanguage":"en"}'

# AI 漫剧转绘(漫画风 720p)
python <SKILL_DIR>/scripts/comic_style.py '{"Vid":"v0d012xxxx","Style":"漫画风","Resolution":"720p"}'

# AI 解说视频(自动生成解说词)
python <SKILL_DIR>/scripts/drama_recap.py '{"Vids":["v023xxx"],"AutoGenerateRecapText":true}'

# AI 剧本还原
python <SKILL_DIR>/scripts/drama_script.py '{"Vids":["v023xxx","v024xxx"]}'

# 查询媒资信息
python <SKILL_DIR>/scripts/get_media_info.py '{"vids":"v001,v002"}'

# 超时后重启编辑类轮询
python <SKILL_DIR>/scripts/poll_vcreative.py <异步智剪任务ID> my_space

# 超时后重启媒体类轮询
python <SKILL_DIR>/scripts/poll_media.py videSuperResolution run_yyy my_space

# 超时后重启翻译轮询
python <SKILL_DIR>/scripts/poll_translation.py <ProjectId> my_space

错误输出

所有错误统一格式:{"error": "说明"}

超时输出(含重启指令):

{
  "error": "轮询超时(360 次 × 5s),任务仍在处理中",
  "resume_hint": {
    "description": "任务尚未完成,可用以下命令重启轮询",
    "command": "python <SKILL_DIR>/scripts/poll_media.py videSuperResolution run_yyy my_space"
  }
}

约束

  • 调用脚本前必须查看脚本详细参数说明

计费说明

仅当用户主动咨询费用或计费规则时,再参考 references/00-billing-instructions.md 中的计费说明,向用户简要说明 byted-vod-process-tools 所依赖的 VOD 资源的计费构成,避免在普通剪辑/处理对话中主动展开计费细节。

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Deploy, update, validate, and troubleshoot the AgentKit hybrid-cloud customer-service demo in this directory. Use when a user asks an AI coding agent to follow the README, deploy or update the demo, configure OpenAPI/Runtime/Knowledge/Memory/Sandbox/MCP/Skills/A2A, run customer-view validation, or record manual platform steps and failures.

日本語の概要は準備中です。原文の説明を表示しています。

bytedance/agentkit-samples4702026年10月9日 更新

Generate deterministic SVG algorithmic artwork. Invoke when the user asks for geometric, generative, or algorithmic visual art.

日本語の概要は準備中です。原文の説明を表示しています。

bytedance/agentkit-samples4702026年10月9日 更新

通过本地 Python CLI 和 OpenAPI 客户端管理、排查火山云手机资源。适用于查询实例和资源、截图、执行命令、查看任务、检查应用、主机和机房容量、标签、DNS、路由,以及操作已授权的测试云手机实例。

日本語の概要は準備中です。原文の説明を表示しています。

bytedance/agentkit-samples4702026年10月9日 更新

Launch and continue the Volcengine AI Research survey workflow for concept testing, audience design, questionnaire drafting, interview guide generation, execution confirmation, progress checks, and result queries. Use this skill when the user wants to create, revise, confirm, execute, or follow up on a real AI research survey task in ABCompass instead of doing generic brainstorming, copywriting, translation, summarization, or broad market discussion.

日本語の概要は準備中です。原文の説明を表示しています。

bytedance/agentkit-samples4702026年10月9日 更新

Create and check long-running video material evaluation tasks. Use this skill when the user wants to submit videos for evaluation, check an existing video evaluation task list, or fetch the result of a previously created video evaluation task. This skill is not a general-purpose video upload skill because upload is allowed only as an internal step of task creation. Authentication uses an API key passed as an Authorization bearer token.

日本語の概要は準備中です。原文の説明を表示しています。

bytedance/agentkit-samples4702026年10月9日 更新

火山引擎 AntiDDoSPro 高防域名只读巡检。用户要查域名健康、攻击、流量、CC/WAF/区域封禁、智能防护或 CCAI 状态时使用。

日本語の概要は準備中です。原文の説明を表示しています。

bytedance/agentkit-samples4702026年10月9日 更新

bytedance のスキルをすべて見る

このスキルの問題を報告する