本文へ移動
cccskills
無料GitHub で公開

media-use

Agent Media OS for a HyperFrames project. Resolve BGM, SFX, image, icon, brand logo, voice, color grade, or LUT into a frozen local file or paste-ready block + ledger record (one verb, `resolve`); generate via TTS / music / image models when the catalog misses; produce voiceover, transcription, captions, and background removal through one shared audio engine; operate on media (cut / reframe / transform); and reuse assets across projects. Also use for vague feedback that real footage looks dark, flat, boring, should feel retro/camcorder/print/ASCII, needs privacy, or needs a media reveal. When the host app provides its own music or sound-effect tools, use those for music and sound effects; `resolve --type bgm|sfx` needs the heygen CLI. When `HEYGEN_API_BASE` is set, HeyGen calls go through that host with no CLI sign-in.

インストール方法を見る

含まれるファイル(110)

  • SKILL.md9.5 KB
  • .gitignore17 B
  • audio/assets/sfx/chime.mp326.7 KB
  • audio/assets/sfx/click-soft.mp311.4 KB
  • audio/assets/sfx/click.mp311.4 KB
  • audio/assets/sfx/CREDITS.md1.2 KB
  • audio/assets/sfx/error.mp350.6 KB
  • audio/assets/sfx/glitch-1.mp382.4 KB
  • audio/assets/sfx/glitch-2.mp3109.5 KB
  • audio/assets/sfx/glitch-3.mp396.8 KB
  • audio/assets/sfx/impact-bass-1.mp366.1 KB
  • audio/assets/sfx/impact-bass-2.mp381.0 KB
  • audio/assets/sfx/key-press.mp33.8 KB
  • audio/assets/sfx/manifest.json3.6 KB
  • audio/assets/sfx/notification.mp376.7 KB
  • audio/assets/sfx/ping.mp325.8 KB
  • audio/assets/sfx/pop.mp322.5 KB
  • audio/assets/sfx/riser.mp3313.5 KB
  • audio/assets/sfx/sparkle.mp356.3 KB
  • audio/assets/sfx/typing.mp326.2 KB
  • audio/assets/sfx/whoosh-cinematic.mp3173.3 KB
  • audio/assets/sfx/whoosh-short.mp318.0 KB
  • audio/assets/sfx/whoosh.mp318.0 KB
  • audio/references/bgm.md8.0 KB
  • audio/references/captions/authoring.md9.3 KB
  • audio/references/captions/motion.md5.5 KB
  • audio/references/captions/transcript-handling.md5.3 KB
  • audio/references/remove-background.md8.3 KB
  • audio/references/requirements.md4.2 KB
  • audio/references/sfx.md3.6 KB
  • audio/references/transcribe.md2.7 KB
  • audio/references/tts-to-captions.md1.7 KB
  • audio/references/tts.md13.7 KB
  • audio/scripts/audio.mjs13.2 KB
  • audio/scripts/audio.test.mjs5.8 KB
  • audio/scripts/gemini-pipeline.test.mjs4.8 KB
  • audio/scripts/heygen-tts.mjs4.3 KB
  • audio/scripts/heygen-tts.test.mjs1.4 KB
  • audio/scripts/heygen-voice.mjs3.8 KB
  • audio/scripts/heygen-voice.test.mjs8.3 KB
  • audio/scripts/lib/audio-meta.mjs1.1 KB
  • audio/scripts/lib/audio-meta.test.mjs4.5 KB
  • audio/scripts/lib/bgm-volume.mjs236 B
  • audio/scripts/lib/bgm.mjs11.3 KB
  • audio/scripts/lib/bgm.test.mjs2.6 KB
  • audio/scripts/lib/concurrency.mjs527 B
  • audio/scripts/lib/concurrency.test.mjs1.6 KB
  • audio/scripts/lib/gemini-auth_test.py4.3 KB
  • audio/scripts/lib/gemini-auth.mjs1.6 KB
  • audio/scripts/lib/gemini-auth.py1.9 KB
  • audio/scripts/lib/gemini-auth.test.mjs3.0 KB
  • audio/scripts/lib/gemini-tts.mjs4.6 KB
  • audio/scripts/lib/gemini-tts.test.mjs8.8 KB
  • audio/scripts/lib/heygen.mjs10.7 KB
  • audio/scripts/lib/heygen.test.mjs10.0 KB
  • audio/scripts/lib/host-audio.mjs1.4 KB
  • audio/scripts/lib/host-audio.test.mjs3.9 KB
  • audio/scripts/lib/main-module.mjs479 B
  • audio/scripts/lib/media-record.mjs3.9 KB
  • audio/scripts/lib/media-record.test.mjs6.9 KB
  • audio/scripts/lib/python.mjs3.2 KB
  • audio/scripts/lib/python.test.mjs5.2 KB
  • audio/scripts/lib/sfx.mjs6.4 KB
  • audio/scripts/lib/sfx.test.mjs4.6 KB
  • audio/scripts/lib/tts.mjs16.8 KB
  • audio/scripts/lib/tts.spawn.test.mjs5.1 KB
  • audio/scripts/lib/tts.test.mjs8.1 KB
  • audio/scripts/lyria-recipe.py4.9 KB
  • audio/scripts/wait-bgm.mjs5.7 KB
  • audio/scripts/wait-bgm.test.mjs3.3 KB
  • luts/index.json2.0 KB
  • luts/README.md1.1 KB
  • references/audio.md2.2 KB
  • references/grading.md6.7 KB
  • references/media-treatment-recipes.md35.0 KB
  • references/media-treatments.md13.8 KB
  • references/memory.md3.4 KB
  • references/meta.md6.6 KB
  • references/operations.md14.5 KB
  • references/resolve.md9.4 KB
  • references/setup-providers.md8.3 KB
  • references/telemetry-dashboard.md4.1 KB
  • scripts/audio-duck.mjs4.0 KB
  • scripts/compatibility.test.mjs3.7 KB
  • scripts/dither.mjs9.7 KB
  • scripts/dither.test.mjs4.6 KB
  • scripts/eval.mjs14.5 KB
  • scripts/lib/config-lock.mjs1.5 KB
  • scripts/lib/cutlist.mjs6.0 KB
  • scripts/lib/duck.mjs3.3 KB
  • scripts/lib/error-diffusion.mjs6.5 KB
  • scripts/lib/index-gen.mjs1.9 KB
  • scripts/lib/manifest.mjs9.1 KB
  • scripts/lib/media-fetch.mjs2.8 KB
  • scripts/lib/media-home.mjs730 B
  • scripts/lib/npx-sync.mjs2.3 KB
  • scripts/lib/parakeet-words.mjs1.1 KB
  • scripts/lib/prefs-store.mjs6.5 KB
  • scripts/lib/recipe-store.mjs12.7 KB
  • scripts/lib/telemetry.mjs6.9 KB
  • scripts/lib/transcriptCutFade.mjs911 B
  • scripts/lib/words.mjs728 B
  • scripts/prefs.mjs2.4 KB
  • scripts/recipe.mjs3.9 KB
  • scripts/resolve-plugin.test.mjs2.5 KB
  • scripts/resolve.mjs1.3 KB
  • scripts/transcribe.mjs7.3 KB
  • scripts/transcribe.test.mjs4.0 KB
  • scripts/transcript-cut.mjs8.9 KB
  • scripts/transcript-cut.test.mjs3.6 KB

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

Plugin installs: Before setup or freshness commands, follow plugin execution rules when this skill is inside a HyperFrames plugin. Standalone installs keep the update instructions below.

media-use

The media OS for HyperFrames: resolve · generate · operate · remember — every media type, one skill, zero context noise.

Only when HEYGEN_API_BASE is set in your environment (a host app set it and pays for HeyGen with its own key): HeyGen media is already paid for. Do not ask the person to install or sign in to the heygen CLI and do not offer its OAuth allowance; catalog search, TTS and avatar calls here go through the host. When that same host also gives you its own HeyGen tools, use those first. When a call through the host is refused, tell the person the host's message as written (it names the fix, such as adding or replacing the key in the app's Settings) and stop; do not switch to another provider unless they ask.

First run otherwise (no HEYGEN_API_BASE), when you will use HeyGen media (catalog search, TTS, avatar video): install and sign in to the heygen CLI (the free-usage path), then verify with npx hyperframes media-use resolve --doctor. Setup and providers: references/setup-providers.md.

Music and sound effects inside a host app: when the app you run in gives you its own music or sound-effect tools, use those. Without HEYGEN_API_BASE, resolve --type bgm and --type sfx search the HeyGen catalog through the heygen CLI; without it they fail and say so (sfx still answers from its bundled library).

Without HEYGEN_API_BASE, before generating a voiceover or an avatar video, tell the person: signing in to the heygen CLI with OAuth (heygen auth login --oauth) gives a free allowance for TTS voiceover and avatar videos, while an API key bills API credits.

Resolve — the one verb

npx hyperframes media-use resolve --type <type> --intent "<description>" --project <dir>

Returns one line: resolved <id> → <path> (<type>, <metadata>). All search noise stays on disk.

TypeOne-line intent
bgmbackground music (HeyGen catalog via the heygen CLI, 10k+ tracks)
sfxsound effects (bundled 19-file library + catalog via the heygen CLI)
imagephotos, backgrounds (HeyGen asset search, 75k+ vectors)
iconicons, symbols (transparent)
logoofficial brand marks (theSVG → GitHub avatar → favicon; never redrawn)
voiceTTS voiceover (HeyGen free-usage path; optional local Kokoro)
grademeasured correction candidate; broad polish/stylization follows Media Treatments
lutuser-provided or explicitly chosen reusable validated .cube file

Before resolving fresh, list reusable candidates with --candidates and judge fit yourself — reuse rules, all flags, ingest (--from), and adopt are in references/resolve.md.

Treat broad visual feedback as media intent

When a user explicitly asks to fix, polish, stylize, obscure, emphasize, or reveal photographic media, read references/media-treatments.md even if they do not name color grading or an effect. Inspect the real <img>/<video>, choose one primary intent, then use deterministic persistence and verification. Use a matching recipe as an optional tested seed, or inspect hyperframes media-treatment --capabilities --json, then request one relevant family/effect with --capability <id> and assemble a custom treatment from canonical controls. Never load --all for ordinary authoring. A treatment may compose correction, a preset, finishing, compatible shader effects, supported keyframes, and optional Registry overlays. Add only source-justified bounded tuning and compatible parts, never effects merely to make the result look more sophisticated. Persist the final combined payload with hyperframes media-treatment.

Use one progressively escalating workflow. For video, inspect one labeled early/middle/late contact sheet rather than reading frames separately. Apply one candidate and inspect one after-sheet for ordinary correction or polish. Escalate to individual frames or moving draft evidence only when the result is ambiguous, temporal, stylized, LUT-based, HDR/LOG-sensitive, private, or brand-critical.

For ordinary correction or polish, persist the final treatment's preset/adjustment JSON. Do not generate a .cube LUT merely to encode exposure, shadows, contrast, or warmth. Use a LUT only when the user supplies one or the selected treatment explicitly owns one. resolve --type grade --for ... --analyze is measurement evidence, not permission to replace the chosen treatment with a generated LUT. Do not recreate supported vignette, grain, blur, pixelate, color, or treatment effects with CSS/SVG overlays; that bypasses Studio controls and the canonical preview/render shader path.

Be proactive — run a media opportunity pass

The human usually can't tell which media would lift the piece. You can. When you build or review a composition, do one grounded scan and then ask once — don't silently add, and don't nag per asset.

Surface an opportunity only when a concrete signal is present:

Signal detectedOffer
On-screen text / a script with no voiceoverTTS voiceover (audio engine)
Emoji or a <div> styled as an iconresolve real icons
Image that is a placeholder, tiny, or upscaled-lookinga better image (and/or upscale — see references/operations.md)
Hard scene cuts / transitions with no soundtransition sfx
A piece over ~10s with no music bedbgm
Footage that reads under/over-exposed or color-casta corrective grade (inspect it with hyperframes media-treatment --selector '#hero' --analyze --json)
Photographic media that feels visually flat or off-topicone specific source-appropriate preset or custom treatment, with the intended target named
A meaningful media entrance/reveal that feels staticone supported seek-safe treatment animation; preserve color unless the request also justifies a preset

Rules that keep this a help, not nagware: grounded, not generic (no signal → no suggestion); opinionated + concrete (propose the specific fix with defaults chosen — the human approves all / some / none); once per project (one consolidated ask; respect "leave it"); surface, never silently mutate (color grades especially: propose and preview — a gray-world "correction" ruins an intentional sunset or neon look).

Where to look — read only the file your task needs

TaskRead
resolve / reuse / adopt / ingest, flags, cascade, inventoryreferences/resolve.md
color grading, LUTs, smart grade (--for), grade-comparereferences/grading.md
voiceover / TTS, music, SFX, captions, transcription (audio engine)references/audio.md
cut / reframe / transform existing media, exact error diffusion, HEVCreferences/operations.md
source-aware creative treatments, realtime effects, overlays, revealsreferences/media-treatments.md
install + auth, provider table, RAM ladders, --local-only, --providerreferences/setup-providers.md
remembered preferences + frozen recipes (user memory)references/memory.md
ownership matrix, usage stats, telemetry, privacy (maintainer-facing)references/meta.md

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Leaves sweep through the frame and part to reveal the headline. HyperFrames block, 1920×1080, 12s, 11 variables.

日本語の概要は準備中です。原文の説明を表示しています。

heygen-com/hyperframes6.1万2026年10月11日 更新

Overlay doctrine for the embedded-captions workflow — the caption MODEL (drop / rail / embed) and the rule that captions are an OVERLAY composited on top of the film, never a reserved bottom band you shift content up to avoid. Load when adding captions/subtitles to a talking-head or launch video, when deciding whether a phrase should be dropped, ride the verbatim rail, or be promoted to a scarce embedded climax, when laying out a composition that will carry captions (do NOT reserve a keep-out band), or when centering a composition on the true frame center under captions. Quotes the rail+embed model from embedded-captions and constraint

日本語の概要は準備中です。原文の説明を表示しています。

heygen-com/hyperframes6.1万2026年10月11日 更新

Turn a weekly changelog .md into a finished branded changelog video (square 1080, ~45-60s, Annie VO, animated brand background, mock-UI visualizations, lowkey captions). Use when the user provides a changelog/digest markdown and wants the weekly video, or says "changelog video". Self-contained — fonts, background, lexicon, and scripts ship in this skill.

日本語の概要は準備中です。原文の説明を表示しています。

heygen-com/hyperframes6.1万2026年10月11日 更新

A tiled headline surface flips cell by cell under a sweeping depth field to reveal the rear headline. HyperFrames block, 1920×1080, 8s, 18 variables.

日本語の概要は準備中です。原文の説明を表示しています。

heygen-com/hyperframes6.1万2026年10月11日 更新

A chain of bevelled cuboids rides a travelling wave as a content carousel. HyperFrames block, 1920×1080, 6.666666666666667s, 40 variables.

日本語の概要は準備中です。原文の説明を表示しています。

heygen-com/hyperframes6.1万2026年10月11日 更新

The technique catalog: five velocity-matched SEAMS (zoom-through, INVERSE zoom-through, cut-the-curve, waterfall cut, rack-focus blur-cut) plus the two in-scene techniques — waterfall ENTRY (staggered arrival cascades for title cards / segment openers) and the nudge curve (slow-fast-slow three-phase group slides). Covers partial-travel (~12% of frame) velocity matching via mirrored power4 eases, the Z scale-sign rule, size-scaled blur (10px text / 18-20px full-frame), word-by-word staggered cuts, cascade pacing by element weight, and the 10/65/25 slide ratio. Read before authoring any transition, text-beat handoff, kinetic text entry, or group reposition. [depth, zoom, inverse-zoom, scale-sign, mirrored-zoom, rack-focus, pacing, velocity, cut-the-curve, waterfall, stagger, cascade, kinetic-text, title-card, segment-opener, nudge, slide, easing, group-motion, z-depth, motion-graphics, cinematic, transition, blur, directional-continuity]

日本語の概要は準備中です。原文の説明を表示しています。

heygen-com/hyperframes6.1万2026年10月11日 更新

heygen-com のスキルをすべて見る

このスキルの問題を報告する