Agent Skills: speech-to-text

Transcribe audio to text with ElevenLabs Scribe and Whisper models via inference.sh CLI. Models: ElevenLabs Scribe v2 (98%+ accuracy, diarization), Fast Whisper Large V3, Whisper V3 Large. Capabilities: transcription, translation, multi-language, timestamps, speaker diarization, audio event tagging. Use for: meeting transcription, subtitles, podcast transcripts, voice notes. Triggers: speech to text, transcription, whisper, audio to text, transcribe audio, voice to text, stt, automatic transcription, subtitles generation, transcribe meeting, audio transcription, whisper ai, elevenlabs stt, scribe, eleven labs transcribe

UncategorizedID: inference-sh/skills/speech-to-text

Install this agent skill to your local

pnpm dlx add-skill https://github.com/inference-sh/skills/speech-to-text

Skill Files

Browse the full folder contents for speech-to-text.

Download Skill

Loading file tree…

Select a file to preview its contents.