Fun, but not enough to be fluent
"Fun, but you will never be fluent unless you work at it with multiple sources… Duo is especially lacking an explanation of the grammar."
Duolingo Redesign
A concept redesign exploring how interleaving, desirable difficulty, and intrinsic motivation could turn Duolingo's habit loop into durable language progress—so retention would follow from real gains rather than a number. A design hypothesis, not a tested result.
The learner sets a pace from a real language level, then works an interleaved daily plan.
01The problem
Duolingo brilliantly rebuilt retention on external motivation—streaks, points, the unhinged owl. This study's premise is that external motivation has limits: in desk-research threads, some learners describe opening the app without feeling closer to fluency, and reaching for tools like italki, LingQ, or Preply. These are self-selected signals, not a measured cause.
"Fun, but you will never be fluent unless you work at it with multiple sources… Duo is especially lacking an explanation of the grammar."
"Many courses teach essentially 0 grammar, which makes learning a language without other tools effectively impossible."
"Courses are stretched out to keep you subscribing… it feels like it's made to keep you on the app rather than reach fluency."
Desk research on r/duolingo and r/languagelearning—a small, self-selected sample. Quotes are paraphrased (not verbatim) and attributed by subreddit: illustrative signal, not a representative measure of all learners.
02Competitor analysis
I compared Duolingo with Babbel and Busuu across how they pace, drill, and motivate—a feature-level read of the three apps in Spring 2025, not an exhaustive benchmark. Three gaps stand out: the first two recur across all three, and the third is where Busuu already does better. Each one sets up a redesign below.
You can't see your level or adjust the challenge.
Lessons crawl through fixed baby steps—there's no way to test in, speed up, or dial difficulty to your real ability.
Skills mix, but items stay easy and repetitive.
Recognize-over-recall keeps challenge below the learner, and error logs surface only recent mistakes—never what you got wrong weeks ago.
Streaks and points, not visible real progress.
The reward is an in-app number, not a rising language level—so when the streak loses its pull, little intrinsic reason remains. Busuu shows it's fixable.
| App | Pace & difficulty | Challenge & review | Motivation |
|---|---|---|---|
| Duolingo | Fixed baby steps | Blocked, recent-only review | Streaks & points only |
| Babbel | Fixed course levels | Blocked-repetitive | Some level picture |
| Busuu | Level test, limited control | Scenario-based, shallow review | Reward tied to real level |
03The cognitive-science lens
Three borrowable principles explain what the gamified habit loop leaves on the table. Each one drives a redesign idea in the next section.
Concept 01 · Contextual interference
Counter-intuitively, mixing tasks and spacing them—so you partly forget between tries—forces the brain to reconstruct knowledge, encoding it more deeply. Blocked repetition feels smoother but fades faster.
Kim et al. (BTF-LTR) · Shea & Morgan (1979) · Schmidt schema theory
Concept 02 · Challenge point
Comfort-Zone Theory and the Zone of Proximal Development describe a sweet spot—hard enough to stretch, not so hard you panic. Challenge Point Theory adds that difficulty must track current skill. Duolingo's fixed baby steps sit below it.
Brown (2008) · Vygotsky, ZPD · Challenge Point Theory
Concept 03 · Self-Determination Theory
Durable motivation comes from autonomy, competence, and relatedness—not points. When learners can see genuine language-level gains and steer their own path, motivation shifts from external streaks to internal purpose.
Deci & Ryan (1985), Self-Determination Theory
Extrinsic
🔥 Streak ◆ Points 🏆 LeaderboardIntrinsic
Perceived real progress04The redesign
Each idea closes one gap from the audit and names the principle it puts to work—so problem, theory, and interface stay coupled.
Gap 01 — no control over pace or difficulty
Redesign 01 · Applies the challenge point
A quick level test places you—"Current level 65"—and a difficulty control lets you speed up or dial in challenge instead of crawling through fixed baby steps. Each option even estimates your next-level date, so difficulty tracks skill and practice stays inside the ZPD.
Gap 02 — blocked practice, and review forgets old knowledge
Redesign 02 · Applies contextual interference
A daily plan interleaves review, comprehension, and dual-task practice instead of blocking them, and a forgetting-curve engine re-surfaces old errors right before you'd lose them. Two tasks fold learning into scenes you already have: "Walk-to-learn" turns a commute into retrieval practice, and "Before-bed Recap" rides the memory-consolidation window during sleep.
Gap 03 — no intrinsic motivation from visible progress
Redesign 03 · Applies Self-Determination Theory
Replace streak-first rewards with a visible language-level track and level-test milestones, so the payoff is perceived competence—real progress up the CEFR scale—not an in-app number. Extrinsic cues remain, but intrinsic progress leads.
05The overall flow
The intended loop: a level test sets your pace; each day interleaves the four learning phases—moving knowledge from working memory into long-term memory; visible level gains are designed to feed intrinsic motivation back into the next loop.
Placement by real ability → Current Level
Daily interleaved practice Working memory → Long-term memory
Review re-surfaces old errors on the forgetting curve · dual-task & real-life prompts push retrieval outside the app
CEFR level up → intrinsic motivation
Overall concept flow, merging the four-phase memory model with the product's navigation.
06Takeaways
Turned cognitive-science literature into concrete interaction decisions—one principle per design move.
Reframed retention as a motivation-and-memory problem, not only a gamification one.
Connected pacing, interleaved practice, spaced review, and visible progress into one coherent loop.
A concept study grounded in desk research and literature.
A research-to-design exploration of applying cognitive science to language-learning UX.