Scribiz
SkillFiles & storageGet the transcript, on-screen text, summary, and chapters of any video or audio with Scribiz. Use when the user wants to transcribe a YouTube, TikTok, Instagram, podcast, or local media file, generate SRT or VTT subtitles, summarize a video, ask questions about a video, or turn a video into a blog post, show notes, or social posts.
Use Scribiz in Claude, ChatGPT or Ahel Desktop
Free. Sign in, add Scribiz and connect your AI. About a minute.
Also: Claude Code · Cursor · Codex
Then ask your AI: use the Scribiz skill
Details
Instructions available. Your AI can read the instructions. Execution depends on the setup they require.
Account requirements not reviewed. Check the skill instructions before use; Ahel provides instructions and does not run this skill.
No other account needed.
Add Ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.
What this skill tells your AI
The instructions your AI receives, as published by illyism/seo-skills in content/scribiz/SKILL.md and read by Ahel’s review.
Scribiz turns a video or audio link, or a local file, into a transcript with speakers and timestamps, the text shown on screen, a summary, and chapters. It works when a video has no captions.
Docs: scribiz.com/docs. Every docs page is also Markdown: add .md to the path.
Pick a way in
What does the agent have?
├── An MCP client (Claude Code, Cursor, Codex) and a link
│ └── Use the MCP server. Best for questions, summaries and search.
├── A shell and a link or local file
│ └── Use the CLI. Best for subtitles, files on disk and batch jobs.
└── Neither
└── Send the user to scribiz.com to paste the link.
MCP server
Hosted at https://scribiz.com/mcp (Streamable HTTP). No key needed to start.
# Claude Code
claude mcp add --transport http scribiz https://scribiz.com/mcp
Cursor, in ~/.cursor/mcp.json:
{ "mcpServers": { "scribiz": { "url": "https://scribiz.com/mcp" } } }
# Codex: ~/.codex/config.toml
[mcp_servers.scribiz]
url = "https://scribiz.com/mcp"
Tools:
get_video_context: summary, chapters and key moments in under 2,000 tokens. Start here.search_video: find where a phrase is said and read only those minutes.ask_video: an answer with 3–5 cited moments, each a timestamp link.get_transcript: the full transcript, when you really need all of it.get_job: check a running job.
Without a key the server reads captions, videos already processed and, within a small daily allowance, the link itself. It never looks at the picture. With an API key (Authorization: Bearer $SCRIBIZ_API_KEY) it uses the account's minutes and can read the screen. Other clients: MCP install.
Prefer get_video_context then search_video over pulling the whole transcript. A two-hour talk is more than 25,000 tokens.
CLI
npm install -g scribiz # or: npx scribiz --help
scribiz login # free account, 30 minutes a month
scribiz setup # or: your own Gemini API key
scribiz doctor # checks the credential, ffmpeg, ffprobe, yt-dlp
Needs Node.js 24+, ffmpeg, and yt-dlp for most links (brew install ffmpeg yt-dlp). Tested on macOS.
scribiz "https://www.youtube.com/watch?v=VIDEO_ID" -o talk.srt # subtitles
scribiz talk.mp4 -f txt -o talk.txt # plain text
scribiz talk.mp4 -f md -o talk.md # Markdown with speakers
scribiz context talk.mp4 --visual -f context # transcript + on-screen notes + summary + chapters, for an LLM
scribiz talk.mp4 --json --out-file talk.json # save the full result once
scribiz format talk.json -f vtt # render another format without paying again
scribiz ask talk.json "What did they decide about pricing?"
Put links that contain ? in quotes. A folder input writes one .srt next to each file.
Options
- Format:
-f srt|vtt|txt|md|context|json(defaultsrt) - Mode:
-m auto|captions|audio|visual|full(aliaseslisten,watch,both). Auto uses good captions and listens when there are none. - Speakers:
--speakersfor interviews and podcasts,--no-speakersfor on-screen captions - Proofread:
--proofread, plus--vocabulary "Brand,Name"for names speech-to-text mishears - Language hint:
-l en - Cost guard:
--max-cost 0.50,--max-duration 02:00:00 - Scripts:
--no-inputnever prompts. Exit codes: 0 ok, 2 usage, 3 sign-in or quota, 4 source, 5 provider, 6 missing tool.
Do not split or rechunk SRT cues after the fact. Re-run Scribiz instead.
Content workflows
- Video to blog post: run
get_video_context(orscribiz context), outline from the chapters, quote the transcript for key claims, and link timestamps as sources. - Podcast show notes:
-f md --speakers, then write a summary, chapter list with timestamps, and pull quotes. - Social clips: use
search_videoto find the strongest moments, then write posts that cite the exact second. - Competitor research: summarize a competitor's demo or webinar and list the claims and features they show on screen (
--visual). - Subtitles for upload:
--proofread --no-speakers -o video.srt.
Limits
- YouTube, podcast feeds, direct media links and files work best. TikTok, X and Vimeo are best effort. Instagram usually needs
--cookies-from-browser chromeon the CLI. - Not for live streams, DRM services (Spotify, Netflix), editing, or generating video.
- Spot-check names and brands in the output before publishing.
Signals
- GitHub stars
- 44
- Forks
- 4
- Last commit
- Oct 2026
Ahel review
K1binfo
installs-packages
Automated review, not a security audit. Ruleset v1+k2.
Advanced
- Item type
- skill
- Key
scribiz- Source
- github.com/illyism/seo-skills
Related picks
Skill · thedaviddias
The pick for JavaScriptmodern-javascript-patterns
Skill · wshobson
The pick for JavaScriptlark-markdown
Skill · larksuite
The pick for Markdownmarkdown-formatter
Skill · nvidia
The pick for Markdownblog-audio
Skill · agricidaniel
The pick for Blogblog-post-interview
Skill · maragudk
The pick for Blog