Comprehensible Input
Comprehensible input is language you understand without decoding it. You read a sentence and the meaning arrives; you do not assemble it. That is the whole definition, and it is the mechanism underneath almost everything on this site.
The term is Stephen Krashen's, from the input hypothesis he set out in the early 1980s. His claim was that people acquire a language by understanding messages in it, and that conscious study of rules produces something different and less useful: knowledge about the language rather than command of it. The formulation usually quoted is i+1 - input pitched a little beyond your current level, where the new material is inferable from the surrounding material you already have.
Krashen's stronger claims have been argued over for forty years and some of them have not aged well. The core observation has aged very well indeed.
The number that actually matters
Strip away the theory and comprehensible input reduces to a coverage figure: what proportion of the words in front of you do you already know?
The research converges on a narrow band:
- 95% is the threshold Laufer identified for reasonable comprehension of a text.
- 98% is the figure Hu and Nation put on comfortable, unassisted comprehension - reading without help and without it feeling like work.
- Below 95%, comprehension collapses fast. You spend your attention decoding, and decoding crowds out the acquisition you are reading for in the first place.
The distance between 95% and 98% is larger than it looks. At 95% coverage you meet an unknown word roughly every twenty words, which is about one per line. At 98% it is one in fifty. The first is study. The second is reading.
This is the number the whole method hangs on, and it has an uncomfortable implication that input purism tends to skate past.
The bottleneck nobody puts on the poster
At zero words, nothing is comprehensible.
That sounds trivial. It is not, because it means the input method is at its weakest exactly where beginners are told to rely on it hardest. The first stretch of pure input is not 95% coverage, it is closer to 5%, and 5% coverage is not language, it is weather. You can sit through a hundred hours of it and acquire remarkably little, because acquisition requires comprehension and there is almost none available.
This is where adults quit. Not because they lack an ear, but because they were told the input would click and it did not click on the timescale anyone promised. The 1,000 to 1,500 hour estimates that circulate in input-only communities are real, and the reason they are so enormous is that a large share of those early hours are spent on material the learner cannot yet understand.
There is a faster way through, and it is not a shortcut. It is arithmetic.
Engineering comprehensibility: the frequency lever
Word frequency in every human language follows a steep curve. A small number of words do most of the work, and the drop-off is severe.
Nation's coverage work on English found that the most frequent 1,000 word families cover roughly 80-84% of spoken discourse. Comparable figures hold across the major languages by Zipf's law. The practical consequence is one of the most useful facts in language learning:
Learning the thousand most frequent words of a language takes you to roughly 80% coverage. Learning a thousand words chosen at random takes you almost nowhere.
Eighty per cent is not yet the 95% threshold. But it is the difference between weather and signal, and crucially it is the point where context starts doing the work for you. At 80% coverage, an unknown word sits in a sentence you otherwise understand, which is precisely the condition under which you can infer it and keep it. Below that, an unknown word sits in a fog.
So the position this site takes is not that comprehensible input is wrong. It is that comprehensibility is a property you engineer rather than a stage you wait for, and the fastest lever is deliberate, frequency-ordered vocabulary work. See why the first 1,000 words matter for the full arithmetic.
The loop
In practice the method is a loop, and the order matters:
- Drill a frequency tier. Top 100, then top 500, then core 1,000, and on. Use active recall, not rereading, and let spaced repetition decide when each word comes back.
- Read and listen inside that tier until it is genuinely easy. This is where the acquisition happens. Graded readers are written to exactly this constraint: every content word sits at or below a frequency ceiling, so a learner who has drilled the top 100 can read a top-100 story end to end with no lookups.
- Move up a tier and repeat.
Each half feeds the other. The drill is what makes the next tier of input comprehensible; the input is what converts drilled words into reflexive recognition and teaches you the grammar you were never explicitly taught. Skip the drill and the input is fog for a very long time. Skip the input and you end up with a large passive vocabulary you cannot use at conversational speed - the classic outcome of a good school languages education.
The Kilo Lingo curriculum is built as this loop. Each chapter drills a band of the frequency list and then hands you a graded story written to that exact band.
Where to get it, by stage
Absolute beginner (0 to ~300 words). Almost nothing authentic is usable yet. Use material engineered for the stage: graded readers at the lowest tier, beginner comprehensible-input video where the speaker draws and gestures, and the word drill to build the floor you are standing on.
Early intermediate (~300 to 1,500 words). The range opens up fast. This is where dedicated CI libraries earn their reputation:
- Spanish: Dreaming Spanish is the reference video library, built explicitly on Krashen's model. Español con Juan is the standard podcast step up. See also the Spanish podcast round-up and the Spanish reading list by CEFR.
- French: InnerFrench is the closest French equivalent and is pitched beautifully at B1. Français Authentique suits slightly earlier. More in the French podcast round-up.
- Mandarin: Lazy Chinese and Comprehensible Chinese cover the beginner video tier. Du Chinese and Mandarin Companion handle graded reading, which matters more in Mandarin than anywhere else because the character barrier makes authentic text inaccessible for far longer. See the Mandarin podcast round-up.
Upper intermediate and beyond (2,000+ words). Authentic material becomes viable, and the job changes from finding comprehensible input to finding input you actually enjoy. Television with target-language subtitles, novels, native podcasts. The coverage threshold is doing the work for you now; your only remaining task is volume.
Four ways people get this wrong
Treating background noise as input. Audio you cannot follow is not comprehensible input, it is audio. Exposure alone does not build vocabulary or grammar. If you cannot follow the gist, the material is not doing the job you think it is.
Choosing material that is far too hard. The instinct is that harder material pulls you up faster. It does the opposite. A text at 80% coverage teaches less per hour than a text at 96%, because at 80% you are decoding rather than reading. Pick the level where you can move forward without stopping.
Choosing material that is too easy for too long. The mirror mistake. Once a tier is comfortable, the new-material rate has dropped and you are maintaining rather than acquiring. Move up.
Believing that vocabulary study is cheating. It is not cheating and it is not opposed to input. It is the fastest available method for reaching the coverage level at which input works at all.
The honest summary
Krashen was right about the mechanism. Language is acquired by understanding things, not by memorising rules about them, and every hour of genuinely comprehensible input is worth several hours of grammar drill.
Where the popular version of the argument goes wrong is in treating comprehensibility as something that arrives on its own if you are patient enough. It arrives much sooner if you go and get it, and the most efficient route is the least glamorous one: learn the most frequent words first, deliberately, then spend as much time as you can reading and listening at the level those words unlock.
That is the entire method. The rest of this site is just an implementation of it.