Skip to content

hopper-inc/hopper

v1.2.3MIT

LLM, text-to-speech and speech-to-text for voice agents and AI agents: set Hopper up in a voice-agent project, measure time to first token on its own prompt, find what slows it down, and speak or transcribe audio.

hopper-transcribe

Transcribe speech in an audio or video file with Hopper speech-to-text, with word timestamps. Takes WAV, MP3, M4A, voice memos and video (decoded with ffmpeg). Needs no setup: the first use registers a key for the agent itself. Use when the user says "transcribe this", "what does this recording say", "speech to text", "STT", "summarize this call" or "caption this". Not for generating audio (hopper-speak) or measuring LLM latency (hopper-benchmark).

Read SKILL.md at the source

Pinned to revision cf8a00a0f3a7, so it is the text this page describes rather than whatever the author pushed since.

Pre-approved tools experimental

Experimental field. Support varies between clients, so this list is what the author declared, not what your client will enforce.

  • Bash(python3 ${CLAUDE_SKILL_DIR}/scripts/hopper_voice.py *)

Files

Every link opens the file at its source, pinned to the revision this page describes.