duo-infra-cost-efficiency
Make infrastructure spend visible, attributable and cuttable without slowing the company down — hourly per-dimension cost tagging, pushing the number into channels engineers already read, defining a win as waste removed rather than spend removed, pricing tech debt in dollars, auditing internal fan-out and cache TTLs, relaxing freshness so caching becomes legal, and deterministic entity-level sampling. Use when someone says our cloud bill is out of control, we cannot tell which team or feature is spending, costs jumped last month and nobody knows why, our LLM spend grows faster than usage, how do we justify infrastructure work to leadership, our analytics queries are slow and expensive, or should we just buy more capacity. For the architectural fix itself, route to duo-backend-architecture.
- Version
- 2.0.0
- License
- MIT
Pinned to revision 78072c9528fb, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/duo-infra-cost-efficiency/SKILL.md
- skills/duo-infra-cost-efficiency/references/audit-the-fan-out-not-the-endpoint.md
- skills/duo-infra-cost-efficiency/references/price-tech-debt-in-dollars.md
- skills/duo-infra-cost-efficiency/references/put-the-number-where-engineers-already-look.md
- skills/duo-infra-cost-efficiency/references/relax-freshness-to-make-caching-legal.md
- skills/duo-infra-cost-efficiency/references/sample-by-entity-not-by-row.md
- skills/duo-infra-cost-efficiency/references/tag-every-dimension-at-hourly-granularity.md
- skills/duo-infra-cost-efficiency/references/target-waste-not-spend.md
Every link opens the file at its source, pinned to the revision this page describes.