Skill · em Criar imagem, vídeo e arte

transcribe

Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.

Procedência

Antes de instalar

7 arquivos · 24,6 KB · inclui 1 script que executa: scripts/transcribe_diarize.py

Instalar na sua CLI

O comando baixa a versão fixada (commit 57f899e) direto da origem, para a pasta que a CLI lê. Precisa de curl (macOS e Linux); no Windows não há comando, porque o Rook Labs é para macOS.

Claude Code

Neste projeto: instala em .claude/skills/transcribe/.

d=".claude/skills/transcribe"
u="https://raw.githubusercontent.com/davila7/claude-code-templates/57f899e5394bb8ca166f38eacae8f0853cbfe033/cli-tool/components/skills/media/transcribe"
curl -fsSL --create-dirs \
  -o "$d/SKILL.md" "$u/SKILL.md" \
  -o "$d/agents/openai.yaml" "$u/agents/openai.yaml" \
  -o "$d/assets/transcribe-small.svg" "$u/assets/transcribe-small.svg" \
  -o "$d/assets/transcribe.png" "$u/assets/transcribe.png" \
  -o "$d/LICENSE.txt" "$u/LICENSE.txt" \
  -o "$d/references/api.md" "$u/references/api.md" \
  -o "$d/scripts/transcribe_diarize.py" "$u/scripts/transcribe_diarize.py"

Global: instala em ~/.claude/skills/transcribe/.

d="$HOME/.claude/skills/transcribe"
u="https://raw.githubusercontent.com/davila7/claude-code-templates/57f899e5394bb8ca166f38eacae8f0853cbfe033/cli-tool/components/skills/media/transcribe"
curl -fsSL --create-dirs \
  -o "$d/SKILL.md" "$u/SKILL.md" \
  -o "$d/agents/openai.yaml" "$u/agents/openai.yaml" \
  -o "$d/assets/transcribe-small.svg" "$u/assets/transcribe-small.svg" \
  -o "$d/assets/transcribe.png" "$u/assets/transcribe.png" \
  -o "$d/LICENSE.txt" "$u/LICENSE.txt" \
  -o "$d/references/api.md" "$u/references/api.md" \
  -o "$d/scripts/transcribe_diarize.py" "$u/scripts/transcribe_diarize.py"

Codex

Neste projeto: instala em .agents/skills/transcribe/.

d=".agents/skills/transcribe"
u="https://raw.githubusercontent.com/davila7/claude-code-templates/57f899e5394bb8ca166f38eacae8f0853cbfe033/cli-tool/components/skills/media/transcribe"
curl -fsSL --create-dirs \
  -o "$d/SKILL.md" "$u/SKILL.md" \
  -o "$d/agents/openai.yaml" "$u/agents/openai.yaml" \
  -o "$d/assets/transcribe-small.svg" "$u/assets/transcribe-small.svg" \
  -o "$d/assets/transcribe.png" "$u/assets/transcribe.png" \
  -o "$d/LICENSE.txt" "$u/LICENSE.txt" \
  -o "$d/references/api.md" "$u/references/api.md" \
  -o "$d/scripts/transcribe_diarize.py" "$u/scripts/transcribe_diarize.py"

Global: instala em ~/.agents/skills/transcribe/.

d="$HOME/.agents/skills/transcribe"
u="https://raw.githubusercontent.com/davila7/claude-code-templates/57f899e5394bb8ca166f38eacae8f0853cbfe033/cli-tool/components/skills/media/transcribe"
curl -fsSL --create-dirs \
  -o "$d/SKILL.md" "$u/SKILL.md" \
  -o "$d/agents/openai.yaml" "$u/agents/openai.yaml" \
  -o "$d/assets/transcribe-small.svg" "$u/assets/transcribe-small.svg" \
  -o "$d/assets/transcribe.png" "$u/assets/transcribe.png" \
  -o "$d/LICENSE.txt" "$u/LICENSE.txt" \
  -o "$d/references/api.md" "$u/references/api.md" \
  -o "$d/scripts/transcribe_diarize.py" "$u/scripts/transcribe_diarize.py"

Antigravity

Neste projeto: instala em .agents/skills/transcribe/.

d=".agents/skills/transcribe"
u="https://raw.githubusercontent.com/davila7/claude-code-templates/57f899e5394bb8ca166f38eacae8f0853cbfe033/cli-tool/components/skills/media/transcribe"
curl -fsSL --create-dirs \
  -o "$d/SKILL.md" "$u/SKILL.md" \
  -o "$d/agents/openai.yaml" "$u/agents/openai.yaml" \
  -o "$d/assets/transcribe-small.svg" "$u/assets/transcribe-small.svg" \
  -o "$d/assets/transcribe.png" "$u/assets/transcribe.png" \
  -o "$d/LICENSE.txt" "$u/LICENSE.txt" \
  -o "$d/references/api.md" "$u/references/api.md" \
  -o "$d/scripts/transcribe_diarize.py" "$u/scripts/transcribe_diarize.py"

Global: instala em ~/.gemini/antigravity-cli/skills/transcribe/.

d="$HOME/.gemini/antigravity-cli/skills/transcribe"
u="https://raw.githubusercontent.com/davila7/claude-code-templates/57f899e5394bb8ca166f38eacae8f0853cbfe033/cli-tool/components/skills/media/transcribe"
curl -fsSL --create-dirs \
  -o "$d/SKILL.md" "$u/SKILL.md" \
  -o "$d/agents/openai.yaml" "$u/agents/openai.yaml" \
  -o "$d/assets/transcribe-small.svg" "$u/assets/transcribe-small.svg" \
  -o "$d/assets/transcribe.png" "$u/assets/transcribe.png" \
  -o "$d/LICENSE.txt" "$u/LICENSE.txt" \
  -o "$d/references/api.md" "$u/references/api.md" \
  -o "$d/scripts/transcribe_diarize.py" "$u/scripts/transcribe_diarize.py"

Peça ao Rook

Já usa o Rook Labs? Cole no chat do Rook: instale a skill https://rooklabs.sh/marketplace/cct.transcribe

Prévia do SKILL.md

---
name: "transcribe"
description: "Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings."
author: openai
---


# Audio Transcribe

Transcribe audio using OpenAI, with optional speaker diarization when requested. Prefer the bundled CLI for deterministic, repeatable runs.

## Workflow
1. Collect inputs: audio file path(s), desired response format (text/json/diarized_json), optional language hint, and any known speaker references.
2. Verify `OPENAI_API_KEY` is set. If missing, ask the user to set it locally (do not ask them to paste the key).
3. Run the bundled `transcribe_diarize.py` CLI with sensible defaults (fast text transcription).
4. Validate the output: transcription quality, speaker labels, and segment boundaries; iterate with a single targeted change if needed.
5. Save outputs under `output/transcribe/` when working in this repo.

## Decision rules
- Default to `gpt-4o-mini-transcribe` with `--response-format text` for fast transcription.
- If the user wants speaker labels or diarization, use `--model gpt-4o-transcribe-diarize --response-format diarized_json`.
- If audio is longer than ~30 seconds, keep `--chunking-strategy auto`.
- Prompting is not supported for `gpt-4o-transcribe-diarize`.

## Output conventions
- Use `output/transcribe/<job-id>/` for evaluation runs.
- Use `--out-dir` for multiple files to avoid overwriting.

## Dependencies (install if missing)
Prefer `uv` for dependency management.

```
uv pip install openai
```
If `uv` is unavailable:
```
python3 -m pip install openai
```

## Environment
…

Ver todo o marketplace