openrouter-benchmarks
Query OpenRouter's Benchmarks API for model benchmark rankings and scores. Use when the user asks for benchmark-backed model selection, model rankings by coding/intelligence/agentic ability, Artificial Analysis or Design Arena ELO/win-rate results, benchmark citations, or wants to call GET /api/v1/benchmarks. Also use alongside openrouter-models when the user asks what model should power an app, product, workflow, or use case and benchmark evidence could inform or rule out part of the recommendation, including creative writing, editing, coding, design, agentic, or intelligence-heavy apps. Do not use for OpenRouter usage analytics, billing/spend analysis, generation metadata, provider uptime/latency, generic model pricing/capability lookup without any selection or benchmark-relevance decision, or creating an evaluation suite for a local app.
Pinned to revision 39b3f1315b59, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/openrouter-benchmarks/SKILL.md
- skills/openrouter-benchmarks/README.md
- skills/openrouter-benchmarks/agents/openai.yaml
- skills/openrouter-benchmarks/references/benchmarks-api.md
Every link opens the file at its source, pinned to the revision this page describes.