Marginal Gains
Part of Mindset · Self-Regulation
Marginal Gains
Marginal gains is a method for choosing what to improve next: the smallest change you are near-certain to pull off, chosen so it builds on the one before it. A small improvement in the wrong place is not a small gain; it compounds with nothing.
The question is never what could I improve. It is what is capping the rest.
What makes a gain a gain
A gain stacks when it adds value on top of existing progress rather than scattering effort. The test is one question: does this next gain add value to the progress I just made, or open an unrelated thread?
Three requirements have to hold together. The improvement targets an important skill that actually moves the goal — chasing a peripheral habit while the bottleneck sits still wastes the compounding, which is why Upgrading Your Dimensions separates short-term foundations from later cognitive growth. It has to persist long enough to find the mistakes, fix them, and let the new way become a habit. And it has to fit the larger operating system, not just score well on a local metric. Miss any one and progress caps. When all three line up, stacking buys two things: growth, and habits. Without the habit, every repetition keeps costing effort and concentration.
One worked case: a first gain in note-taking — compressing difficult concepts into fewer keywords — improved retention and stayed slow. Of the two candidates for next, strengthening the same deep processing stacked; reducing procrastination helped and built on nothing. The operating layer that walks that test in short form is Marginal Gains in Practice.
Poor stacking looks like a switching spiral: non-linear notes, then flashcards, then a method for reducing procrastination, then time management, then back to note-taking, unimproved. Weeks or months pass. Divided effort slows acquisition because each skill needs enough experience and reflection to make mistakes, understand them, and build each increment to a meaningful level. Mistakes surface more slowly, experiments lose focus, and each skill takes longer. Weeks pass with surface familiarity across many areas and competence in none — the plateau The Technique Is Only as Good as the Thinking It Produces names, where better methods leave results unchanged. A clean acquisition creates headroom. Switching creates work without creating competence.
Two fronts is the working number. Three is the ceiling. More than three overwhelms.
Effort without progress is not one failure. Fluctuating gains come from inconsistency — real increments that do not build on each other — and are repaired by stacking. Negative gains come from unreliable information, calibration, or feedback, and are repaired by getting feedback, not by persisting harder. From inside they feel identical. Persisting harder on a negative gain makes it worse.
Whether a gain will compound at all is a removal question, not a size question: if the input vanished, would you still be more capable? That test lives in Compounding vs Additive Gains.
Where the gain has to sit
A rate limiter is the part of a process that prevents every other part from improving. Fixing anything else while it remains does not raise the ceiling. Picture a bucket with a hole in its side: perfect everywhere else and it still will not fill above the hole. Work on the handle or the rim changes nothing. The water you pour in meanwhile is wasted.
The same principle one level up is Dimensions of Learning — five capabilities, and the weakest sets the ceiling for the rest.
Limiters are not always obvious. A wrong guess still beats not looking: attempting the wrong limiter still tends to produce smoother progress than charging ahead without looking for one. The discrimination that keeps the hunt from becoming a complaint is: most consistently constraining, not most recently frustrating.
The first place to look is the enablers. Self-management — procrastination, time management, prioritisation, focus — and the growth skills, experimentation and critical reflection, limit execution across the board. A strong learner can be capped by heavy procrastination alone. Old cue-response habits that fire faster than a new skill can stabilise are a candidate limiter of their own, treated in How to Unlearn Old or Bad Habits Efficiently.
Limiters rotate. Limited by a Fixed vs Growth Mindset early, a learner may be limited by time management a month later — because time management got worse, or because the mindset improved. When a constraint breaks, return to the hunt and do not let inertia keep you optimising the old one.
The check runs every one to two weeks: is there a part of my process that seems to be holding everything else back? If yes, that is the next target, ahead of other improvement work. You may already be working on it, in which case continue. A plateau is the same trigger arriving as an event rather than a calendar date.
A technique can be improved forever, so “good enough” cannot come from the skill. It comes from position: you work a skill while it caps the rest, and you move when it stops being the cap.
One selection rule, then you run it
Does it build on the last one, and is it on the thing capping the rest? Two questions, one selection rule. Selection is half the method. The other half is running it.
Direction lives here. The improving lives in Kolbs Experiential Cycle — the four-stage loop that turns one real attempt into a better next one. This without the cycle becomes wishful planning. The cycle without this becomes unfocused reflection — rumination. You cannot run a reflection cycle on every problem every day, so something has to choose which problems get one.
The 30-Day Plan describes what to do for the next thirty days. This page describes how to improve inside that plan. Combined, the order is: name the goal, then the performance goals, then the habits and environment in the way; then two or three skills tracked and stacked; then pace adjustments.
| Aspect | Marginal Gains | 30-Day Plan |
|---|---|---|
| Primary focus | Skill development (getting better at a skill) | Self-management (taking action productively) |
| Core insight | Skills improve faster and more sustainably through compounding tiny gains than through large jumps | Goals are more likely to be achieved when plans account for barriers in advance — especially around habits and tendencies |
| Method | Tracking and stacking 1% gains through experimentation and reflection | Clear goal-setting and habit review to create a high-probability plan |
| Usage | Continuously, wherever skill development is desired | When a concrete, structured plan is needed to make rapid progress toward a specific goal |
| Scope | Breaks bigger skills into smaller focused gains | Breaks bigger goals into 30-day mini-goals with a clear plan |
| End goal | High competence in the target skill | Achievement of the medium-term goal |
The loop, run forward
REDO is the sequence that turns a marginal gain into a stacked improvement. It runs after each serious attempt, looking forward, and it ends in a plan.
R — Reflect. What have the previous attempts taught, looking for trends across multiple attempts, and what might the next increment be? One attempt cannot show a trend.
E — Evaluate. Which direction. The fork is: unblock yourself by moving to another skill, or unlock this one with more experimentation, theory, and reflection.
D — Define. Name the increment, specific enough to test, small enough to stay with — and estimate how long results will take. Early improvement usually looks like finding mistakes rather than performing better. One worked case allowed at least ten hours across a week to surface the major mistakes, then another week or two to test fixes.
O — Optimise. Make a specific plan, run it, reflect again from step one. The goal is data, not perfection.
You will not see improvement on every attempt. What you should see is a trend toward clarity and awareness, with performance later. That is the checkable expectation. No attempt, nothing to reflect on — a house guard, not the definition of the loop. Not knowing how to start is a finding: the next move is a best attempt, because working out how to start can run forever.
Run a cycle after each serious attempt. The form you actually open is Kolbs Template.
Picking the 1%, and being able to see it
The five steps, used once to set the field and then kept light:
- Anchor a meaningful goal, nine months to three years out. Outcomes are controlled indirectly through the skills, attributes, habits, resources and actions that raise their probability — the mechanism in Reverse Goal Setting.
- Dissect it: what knowledge, attributes, skills, processes, resources does it need — then ask how sure you are. You do not need the full list, only the parts clearly necessary as the next step. At the very beginning, the step is sometimes just taking a step. The structured walk that produces this inventory is Skills Audit, run after rushing, after a break, or at a plateau.
- Rate the target level out of ten and say in words what that number means.
- Rate the current level out of ten and justify it as objectively as you can.
- Name the increment that closes the gap.
The justification names the gain. No knowledge → go learn. Knowledge but no practice → run one real attempt. Knows how but has never done it → do it once. Later reviews repeat only steps 4 and 5, because the first three barely change.
When you cannot see a gain, information is the gain. You cannot take action before knowing what action to take. Know enough to attempt it and start making mistakes → experiment. Not enough even to make a mistake → go learn first. Gains progress knowledge → actions → behaviours and habits → position, meaning how much control you have over reaching the goal.
You cannot feel a 1% change. The brain registers small gradual movement badly, so early on there is no honest intuition about whether you are gaining, which is why invisible progress is one of the most common causes of early demotivation. Stacking keeps the direction positive; tracking keeps it visible. The artefact is one notebook: the targets you are on, a written win criterion that tells a gain from a loss, reviewed weekly. That is not a second dashboard. The gain should make the next attempt easier, cleaner, or more motivating; the day the record becomes a dashboard it has stopped being a gain.
The same logic runs outside studying. In fast-moving work, small improvements to search, review, source hygiene, feedback speed and sense-making loops compound into a stronger operating system — the claim in Nothing Ever Happens Is Over. A skill gain pays out everywhere the skill is used.
Keep the increment small enough to start and meaningful enough to repeat. Large strides trade safety for an unlikely boost.
For Priority 0+1 — one to three priorities for tomorrow, two the usual working number — a gain should usually improve one of five things: recurrence, emotional engagement, identity reinforcement, visible progress, action initiation. Bare operators in the table are defined where they first appear as pages: BHS, SIR, Shortcut detection, and the approach-layer work in Interleaving for Complex Problem Solving.
| Priority 0 Area | Skill Or Process | Possible Marginal Gain |
|---|---|---|
| Agentic Engineering | Prompting agents | Save one reusable prompt pattern after it works. |
| Agentic Engineering | Agent workflow design | Add one clearer instruction to AGENTS.md or a project README. |
| Agentic Engineering | Code review with agents | Ask for one focused review category instead of a broad review. |
| Agentic Engineering | Verification | Add one repeatable build, test, or screenshot check to the workflow. |
| Learning Systems | BHS | Improve the Aim step by writing sharper why/how questions before reading. |
| Learning Systems | SIR | Add one interleaved prompt that forces comparison instead of recognition. |
| Learning Systems | Complex problem solving | Change one meaningful variable before executing again. |
| Learning Systems | Kolbs | Reflect on one real attempt instead of journaling about the whole system. |
| Learning Systems | Shortcut detection | Name one shortcut and add one constraint that makes it harder to repeat. |
| Vietnamese | Immersion recurrence | Pick one default YouTube channel or playlist for low-friction starts. |
| Vietnamese | Noticing | Capture one repeated phrase or grammar pattern from real input. |
| Vietnamese | Listening | Rewatch one short clip until it feels less noisy. |
| Vietnamese | Comprehension | Adjust subtitles, speed, topic, or lookup tools to keep attention engaged. |
| 中文 | Maintenance | Read or convert one useful sentence without reactivating the whole language track. |
| 中文 | Character contact | Review one character, phrase, or short clip. |
| Fitness | Strength | Add one small progression: weight, rep, set, tempo, or form cue. |
| Fitness | Mobility | Choose one mobility bottleneck and repeat a short routine. |
| Fitness | Cardio | Make the start easier: shoes ready, route chosen, timer preset. |
| Fitness | Recovery | Track one signal: sleep, soreness, energy, or readiness. |
| Relationships | Contact | Send one message that would otherwise stay vague. |
| Relationships | Repair | Clarify one small tension before it becomes stale. |
| Relationships | Appreciation | Express one specific appreciation. |
| Relationships | Planning | Put one next touchpoint on the calendar. |
Five questions, and a discard rule. Does it reduce start friction? Does it make the skill feel more alive? Does it reinforce the identity? Does it leave visible progress? Does it improve the next attempt? A gain that scores on none of them is organisation work, not improvement work.
Position over size
Position beats size, now that the two questions have been paid for. At this page’s own one-to-two week review cadence, a compounding 1% is about 1.7× over a year. The famous 37× needs a daily cadence and every gain compounding — a cadence this page never recommends.
Large strides cost three things: high failure risk, more variables and a smaller margin; low iteration frequency, feedback arriving in weeks or months instead of days; and misalignment risk. Rebuilding a whole workflow around a tool, then finding that the methods that matter are impossible in it, is the misalignment in one picture.
It should feel like narrowing the improvement field: you leave with one small, believable upgrade that connects to a larger goal. Good signs — specific enough to test, on a real bottleneck, repeatable across sessions, makes progress concrete. It has become avoidance when you keep planning tiny improvements and never run the next attempt. Not knowing how to start is the same finding in other words: make a best attempt.
The gain also improves the picking. Learning your own tendencies both prevents the next mistake and makes the next reflection more accurate.
Related
- Mindset — parent dimension; the material the fixed-mindset limiter example depends on.
- Self-Regulation — second parent; the page is dual-parented.
Open Questions
Early cycles that buy clarity rather than performance, and fluctuating gains that never compound, feel identical from inside — effort without visible progress — and they take opposite repairs. How early can the two be told apart?
Sources
- Theory of Constraints — the public name for the rate limiter, and the five focusing steps whose fifth supplies the inertia warning after a constraint breaks. https://en.wikipedia.org/wiki/Theory_of_constraints
- Compounding arithmetic, stated as computation rather than as a citation: 1.01^52 ≈ 1.68; 1.01^365 ≈ 37.8; 0.99^365 ≈ 0.026.