本文へ移動
cccskills
無料GitHub で公開

transcribe

Transcribe audio files locally with Interpreter's builtin-transcribe tool server. Use when a user asks to transcribe speech from recordings without cloud transcription or API keys.

インストール方法を見る

含まれるファイル(6)

  • SKILL.md2.7 KB
  • agents/openai.yaml386 B
  • assets/transcribe-small.svg750 B
  • assets/transcribe.png1.3 KB
  • LICENSE.txt10.5 KB
  • references/api.md780 B

SKILL.md(原文)

インストールする前に、エージェントに与えられる指示の中身を確認できます。

Local Audio Transcribe

Transcribe audio locally through Interpreter's app-side builtin-transcribe tool server. Do not install Python packages and do not ask for an OpenAI API key. The agent asks Interpreter to download/run local Whisper models; the app server owns the user-data install location and filesystem permissions.

Workflow

  1. Collect the audio/video file path and ask whether the user wants a fast, small model or a larger, higher-quality model when that tradeoff matters.
  2. Run interpreter-app tools builtin-transcribe list_transcription_models --json '{}' to check installed models and compare size/language/quality.
  3. If the chosen model is not installed, ask the user before downloading it. Then run interpreter-app tools builtin-transcribe download_model --json '{"model":"tiny.en"}' or the selected model ID. Interpreter downloads into app user data and shows a download toast.
  4. Run interpreter-app tools builtin-transcribe transcribe_audio --json '{"audioPath":"path/to/audio.mp3","model":"tiny.en","outputPath":"output/transcribe/transcript.txt"}'.
  5. Validate the transcript and save outputs under output/transcribe/ when working in this repo unless the user requested another path.

Decision rules

  • Default to tiny.en for quick English transcription.
  • Use tiny for quick multilingual transcription.
  • Use small.en or small when accuracy matters and a larger download is acceptable.
  • Use medium.en, medium, or large-v3 only when the user accepts much larger downloads and slower local runtime.
  • If transcribe_audio says the model is missing, do not work around it in shell. Ask the user whether to download a model, then call download_model.
  • Speaker diarization and known-speaker labeling are not supported by the local builtin-transcribe tool.
  • For video files, first extract audio with a normal local media workflow, then pass the audio file to transcribe_audio.

Output conventions

  • Use output/transcribe/<job-id>/ for evaluation runs.
  • Use outputPath for transcript files to avoid overwriting.

Tool quick start

interpreter-app tools builtin-transcribe list_transcription_models --json '{}'
interpreter-app tools builtin-transcribe download_model --json '{"model":"tiny.en"}'
interpreter-app tools builtin-transcribe transcribe_audio --json '{"audioPath":"interview.mp3","model":"tiny.en","outputPath":"output/transcribe/interview.txt"}'

Reference map

  • references/api.md: local model IDs and supported audio format notes.

レビュー

まだレビューはありません。使ってみた感想をお寄せください。

同じリポジトリのスキル

概要と使いどころ

Build a period-end accrual schedule by computing each accrual, citing support, and drafting journal entries for controller approval. Use during month-end close; this drafts entries only and must not post them.

日本語の概要は準備中です。原文の説明を表示しています。

openinterpreter/interpreter-workstation232026年10月11日 更新

audit-xls

無料

Audit a spreadsheet for formula accuracy, errors, and common financial-model mistakes. Use for selected ranges, single sheets, or whole-workbook model checks including balance-sheet balance, cash tie-out, roll-forwards, and logic sanity.

日本語の概要は準備中です。原文の説明を表示しています。

openinterpreter/interpreter-workstation232026年10月11日 更新

Use this skill when the user needs advanced Playwright control of an already running browser session through Interpreter's app-managed browser bridge and the `builtin-js-repl` `js_repl` tool, after simple browser page work cannot be handled by the unified `builtin-interpreter` browser page tools.

日本語の概要は準備中です。原文の説明を表示しています。

openinterpreter/interpreter-workstation232026年10月11日 更新

Drive a native desktop app through Interpreter's builtin-cua-driver. Use get_app_state, click/type/scroll/drag, and verify by calling get_app_state again when the user asks to operate a real desktop app, browser chrome, native dialog, menu, secure prompt, file chooser, or hidden/background window.

日本語の概要は準備中です。原文の説明を表示しています。

openinterpreter/interpreter-workstation232026年10月11日 更新

doc

無料

Create, inspect, and edit Word documents (`.docx`) with local code-execution tools. Use for document drafting, structured edits, tables, comments, and layout-sensitive Word deliverables.

日本語の概要は準備中です。原文の説明を表示しています。

openinterpreter/interpreter-workstation232026年10月11日 更新

Design or revise Interpreter UI for non-developers with an ultra-minimal, utilitarian style inspired by the OpenAI website and ChatGPT, not the developer platform. Use when simplifying settings, onboarding, chat, forms, empty states, or helper copy; when removing extra containers, labels, or visual noise; when replacing technical wording with plain language; when clarifying hierarchy and making the main user action obvious; when making every visible label and status understandable to an average HR admin or other non-technical user; when hiding implementation details such as logs, IDs, endpoints, paths, or protocol names; or when a screen must respect the app's existing CSS variables and light/dark theme behavior.

日本語の概要は準備中です。原文の説明を表示しています。

openinterpreter/interpreter-workstation232026年10月11日 更新

openinterpreter のスキルをすべて見る

このスキルの問題を報告する