Skip to content

hktitan/duolingo

v2.0.0MIT

The Duolingo playbook as agent skills — retention, streaks, gamification, learning science, curriculum, efficacy measurement, metrics, experimentation, growth, brand, voice, culture, hiring, platform and mobile engineering, observability, and LLM feature design. Distilled from 750 posts on blog.duolingo.com plus the Duolingo Handbook and Design System. Start at the /duolingo router; design and UI craft routes out to design-engineering.

duo-measurement-validity

Design and report a number that survives an outsider's scrutiny — prove the outcome on an instrument you did not build, break a headline score into components only when they add information, publish the component where you lose, audit bias at intersections instead of one variable at a time, and disclose exactly where your own hand touched the evidence. Also covers the inference traps that make an instrument lie — classifier categories that encode their training population, resemblance mistaken for lineage, non-independent samples, contested definitions. Use when someone asks will this number survive outside scrutiny, is our metric actually valid, should we ship a subscore, why doesn't our completion rate convince buyers, how do we prove our product caused this, is our model biased, can we trust this benchmark, or how do we measure something nobody can measure directly.

Version
2.0.0
License
MIT
Read SKILL.md at the source

Pinned to revision 78072c9528fb, so it is the text this page describes rather than whatever the author pushed since.

Files

Every link opens the file at its source, pinned to the revision this page describes.