hopper-benchmark
Measure a voice agent's LLM time to first token on Hopper with its own system prompt and tools over a simulated 10-turn call, and report first-turn and later-turn latency and the prompt-cache hit rate. Changes no code. Use when the user asks "how fast would my agent be on Hopper", "benchmark TTFT", "is prompt caching working" or wants before/after numbers. Not for STT/TTS latency or load tests; to switch the agent to Hopper, use hopper-integrate.
Pinned to revision cf8a00a0f3a7, so it is the text this page describes rather than whatever the author pushed since.
Pre-approved tools experimental
Experimental field. Support varies between clients, so this list is what the author declared, not what your client will enforce.
- Bash(python3 ${CLAUDE_SKILL_DIR}/scripts/hopper_trial.py)
- Bash(python ${CLAUDE_SKILL_DIR}/scripts/hopper_ttft.py *)
- Bash(python3 ${CLAUDE_SKILL_DIR}/scripts/hopper_ttft.py *)
Files
- skills/hopper-benchmark/SKILL.md
- skills/hopper-benchmark/agents/openai.yaml
- skills/hopper-benchmark/scripts/hopper_trial.py
- skills/hopper-benchmark/scripts/hopper_ttft.py
Every link opens the file at its source, pinned to the revision this page describes.