Two delivery partners on what an eleven-million-dollar education bond actually asked of the classroom.
The Quality Education India Development Impact Bond ran from 2018 to 2022, paid against measured learning outcomes, and routed eleven million dollars through four NGOs into the hands of two-hundred-thousand primary-school children. Pankaj Jain built Gyan Shala in 1999 around the realisation that an ordinary village teacher cannot be expected to perform like a college professor — and that the school system has to be designed around the teacher who actually exists. Prem co-built Pratham InfoTech in 1998 around a donation of two hundred used 486 computers and the belief that a hybrid of technology and tradition is the only model that survives Indian scale. This conversation is what those two playbooks looked like when paid for on results.
In sixty seconds.
India's school-education problem is no longer access. The ASER number Prem reads off the page is 98.5%: that is enrolment. The companion number — also from ASER — is that fewer than half the children in Class 5 in a government school can read a Class 2 text. Enrolment is solved. Reading is not. The whole conversation sits inside that gap.
Pankaj Jain's answer is structural. Gyan Shala does not try to upskill the teacher into a college professor. It accepts the village 12th-pass woman as the teacher who actually exists, and engineers the classroom — three kinds of interaction, structured pedagogy, written-friendly furniture, three to four hours, never more — around her capacity. The yield against a low-cost private-school comparator group ran two to three times above the contracted target across four years.
Prem's answer is hybrid. Two hundred used 486 computers in 1998 became, by 2019, fifty-five Lucknow government schools equipped with twelve-hundred Chromebooks running MindSpark — an adaptive-learning platform built with Educational Initiative. The Class 5 child reading at Class 2 logs in at Class 2; she does not stay lost in a Class 5 lesson. The treatment group rose from a nine-scale-point baseline to forty-three to forty-five. A five-fold lift, audited by a third party, was the bond's currency. Neither delivery partner thought the DIB redesigned their pedagogy. Both said it formalised their monitoring and forced the testing to be done by someone they did not pick.
Where to land in the conversation.
Each chapter opens YouTube at that timestamp in a new tab.
Six ideas to carry into your own work.
The structural arguments inside the conversation, lifted out so they travel. Each one is the kind of thing you can quote in a programme review on Tuesday.
Design for the teacher you actually have.
Pankaj's plainest line. Most school reform programmes assume the teacher will rise to meet the curriculum. Gyan Shala starts from the opposite premise: a 12th-pass village woman with limited training is the teacher who shows up, and the entire classroom — script, materials, timing, the three kinds of interaction — has to be engineered around her capability. Structured pedagogy is the operational name for that decision.
Pleasant, not fun.
The classroom that retains a child is not the one that entertains her. It is the one in which the pleasure of empowerment — the unmistakeable feeling of having learnt something — happens reliably. Fun is a category mistake; it makes "fun" the objective and learning the by-product. Pankaj reverses the dependency. The class must lead the child to the experience of learning; from that experience the willingness to come back follows.
The three classroom interactions.
Inside any class period, exactly three interactions happen: child with teacher, child with other children, child with the learning material. Most reform programmes obsess about the first, occasionally upgrade the third, and ignore the second. Gyan Shala designs all three on purpose, in the order they affect outcome. The classroom is not a relationship — it is a network with three edges, each of which has to be loaded.
Teaching at the right level.
Prem's adaptive-learning answer to the ASER gap. A child in Class 5 who reads at Class 2 will not learn from a Class 5 lesson; she will dissociate. MindSpark routes her into a Class 2 path, and lifts her one rung at a time. The diagnostic and the placement are the program. The lesson is just the consequence. It is the only mechanism the conversation describes that works at scale without adding teachers.
Outcomes-based finance, viewed from the delivery side.
The DIB did not redesign Gyan Shala's pedagogy or Pratham InfoTech's hybrid model. What it did, in both their words, was three things: (i) lock the contractual learning outcomes to the state curriculum's expected outputs, (ii) formalise process monitoring that was previously informal, and (iii) hand the testing to a third party with its own framework. The first re-phased the curriculum slightly. The second cost time. The third required negotiation. None of them changed what worked.
Money is not the constraint. Result is.
Pankaj's flat correction of the Bangladesh comparison. India has spent more per capita on education than Bangladesh for decades, and Bangladesh moved its girl-child cohort to per-capita parity twenty-five years ago. The constraint was never the budget; it was the inability to convert a budget into an outcome. Outcomes-based finance is, in that frame, a structural answer to the right problem.
Fifteen things to actually walk away with.
Each one carries the timestamps where the moment lives, and a transferable note for work that isn't education.
The teacher capability question is the design question.
Pankaj's argument is sharper than the standard NGO line about "the teacher problem". He says explicitly that mainstream education thinkers have spent twenty years opposing structured pedagogy on the grounds that it de-skills the teacher. His response: the average primary-school teacher in rural Gujarat is not being de-skilled; she is being respected. The system being asked of her — to behave like a four-year-trained college professor managing a forty-child mixed-ability classroom — is the de-skilling, because it is impossible.
The Gyan Shala design starts from the actual skill level and works backwards. Lesson scripts are pre-structured. Materials are pre-printed. The three interactions are pre-engineered. The teacher does not need to be the curriculum designer; she needs to be a competent executor of a curriculum designed by someone who understands how a child's mind acquires language, deductive thought, and inductive thought.
Three to four hours of school is enough, if the three to four hours are right.
The unusual claim. Pankaj says a child can achieve mastery of grade-level learning in three to four hours a day, in a classroom that is non-threatening and non-paralysing, with the three interactions running properly. Gyan Shala runs in that frame: a focused supplementary or parallel session that the slum child can fit alongside whatever else her day demands.
The implication for policy is uncomfortable. Indian government schools run six to seven hours; private schools run longer. If three to four well-designed hours match or exceed the outcomes of six to seven undesigned hours, the question is not how much instruction time exists but how it is being used. The conversation does not labour the point. It does not have to.
Drop-out is mostly a value problem, not a poverty problem.
Vishal's framing question was the standard one: poor families pull children out of school because they need them to earn. Pankaj reframes it. In his read of Gyan Shala's experience, ninety percent of dropouts are not earning-pressure — they are children who concluded the school is not giving them enough to stay. The five-to-ten percent who drop out for economic or social reasons are real, but they are the minority case.
The implication is that drop-out is downstream of learning outcomes, not upstream of them. Pay families to keep children in unhelpful schools and the children still leave; build a school in which children visibly learn and the families do not need paying. The lever is the classroom, not the cash transfer.
"Pleasant," and the careful avoidance of "fun."
The most precise small moment in the conversation. Vishal asks how to make school fun; Pankaj declines the word. Fun risks becoming the objective, and an objective of fun is satisfied by sweets and games, not by learning. Pleasant is the right frame, and the pleasure that retains children is specifically the pleasure of mastery — the unmistakable internal signal that you have learnt something.
The distinction is not pedantic. Twenty years of edtech in India has been built on the fun frame and largely failed. Prem says explicitly later in the conversation that the assessment-and-grading edtech company he saw fail "very weird man" was operating in that frame. Pleasure-of-mastery is what carries a child back into the room the next day.
The classroom is a three-edge network, not a relationship.
This is the most operational moment in Pankaj's section. Inside a class period, exactly three interactions happen: child with teacher, child with other children, child with the learning material. He calls them out as a designable triple and notes that all three have to be engineered in coordination, not in isolation.
Most reform programmes obsess about the first edge (teacher-training initiatives), occasionally touch the third (better textbooks), and almost never touch the second (peer interaction). Gyan Shala's classroom design is unusual in that it loads all three on purpose, with the second edge contributing a non-trivial share of the outcome. The Class 5 child does not learn only from the teacher; she learns from the Class 5 child next to her, in a structured peer arrangement that the curriculum designer has set up.
Mind, not skill: how learning is actually modelled.
Pankaj pushes back when Vishal compresses his model into "deductive learning". His correction is patient. Language is its own thing — children acquire language; they do not learn it the way they learn mathematics. Mathematics is most cleanly the home of deductive thinking. Science is the home of inductive thinking. The classroom designer has to understand these three different ways the human mind engages with knowledge, and design class processes accordingly.
The seemingly academic distinction is structural. A class that teaches mathematics with the cadence of a language class will fail; a class that teaches a language with the cadence of a science class will fail differently. Gyan Shala's curriculum is built on the recognition that these are three different operations on the mind, and that an effective classroom switches between them deliberately.
ASER 98.5%: the access war is won. The learning war has not started.
Prem's number is exact. The Annual Status of Education Report — Pratham Education Foundation's flagship survey, running for fifteen years at the time of recording — puts primary-school enrolment at 98.5%. India has solved the access problem that consumed two decades of policy attention. The companion number, also from ASER, is that fewer than half of Class 5 children in a government school can read a Class 2 text. Five years of schooling produces a child who is, in literacy terms, three years behind.
The diagnostic implication is that primary education in India is no longer an enrolment problem and is now a learning-outcomes problem. The two require different instruments. Enrolment was moved by mid-day meals, school construction, RTE-style entitlements. Learning will be moved, if at all, by classroom design, adaptive placement, and outcomes-based monitoring. The DIB is a first instrument calibrated for the second war.
The DIB is a financing instrument, not a pedagogy.
Both Pankaj and Prem refuse the leading question. Vishal opens by asking how the Quality Education India DIB changed the way they teach. Pankaj corrects: a bond is a financing mechanism. It has nothing to do with the education per se. What worked in a classroom worked whether it was funded by a DIB, a grant, or a government cheque; the design of the learning process did not change because the funding structure did.
What did change was governance. The DIB locked the contractual learning outcomes to a negotiated subset of the state curriculum, formalised process monitoring that had previously been informal, and required testing to be done by a third party with its own framework. None of this is pedagogy. All of it is overhead, sometimes useful, sometimes costly. The intellectual honesty here is rare — outcomes-based finance is often sold as a pedagogy transformer, and it is not.
The comparator group revealed who held the system to account.
An overlooked detail. The DIB structure required Gyan Shala's progress to be benchmarked against a comparison group of similar socioeconomic profile. The ideal comparator would have been a government school. The government — Pankaj says with diplomatic understatement — would not agree to be the benchmark. The comparator that was negotiated was a set of low-cost private schools serving children of broadly similar background.
This is two findings, not one. First, the government's reluctance is itself evidence of where accountability for outcomes sits in the Indian primary-education system: nowhere stable. Second, the comparator that was chosen is the harder one — low-cost private schools are paid for by aspirational parents and have stronger incentives than government schools to produce visible literacy. Gyan Shala outperforming the comparator by two-to-three-times the contracted target is therefore the stronger version of the result, not the weaker one.
The test became a test of comprehension, not literacy.
A second overlooked detail of the bond's design. The independent tester chosen by the British Asian Trust did not run a literacy test (read these letters, decode these sounds). It ran a comprehension test (read this passage, understand it, answer questions on it). Pankaj is clear that this is a stronger test, and that the Class 5 children failing the ASER literacy test would have failed this one more severely.
The choice mattered. A test of comprehension forces the curriculum to teach for comprehension; a test of literacy permits the curriculum to teach for decoding. Gyan Shala's outperformance on the comprehension test is a stronger claim than the same outperformance on a literacy test would have been, because comprehension is harder to game with surface drilling. The DIB structure, by handing the testing to a third party with sophisticated instruments, raised the bar for the delivery partner and made the result more credible.
Lucknow, Mission 35, and the daily-data discipline.
Pratham InfoTech's Lucknow build is the most operationally specific stretch of the conversation. Fifty-five Uttar Pradesh government schools, twelve hundred Chromebooks, a learning lab per school with twenty-to-twenty-five computers in one classroom, MindSpark adaptive software underneath. Fifty-five further schools chosen as a matched control. The treatment was a daily target — thirty-five minutes of MindSpark per child per day. They called it "Mission 35." A schedule that took weeks of negotiation with principals and management to fit into the school day.
The daily-data discipline was the hidden lever. Because MindSpark logged every interaction, Pratham could rank schools by daily performance, identify a top five and a bottom five, and propagate the practices of the top five to the rest on the same day's cycle. That feedback loop was tighter than anything paper-based monitoring could have produced. The treatment group ended up at a forty-three-to-forty-five-scale-point baseline versus the control at nine. The fivefold lift is the headline; the daily ranking-and-propagation is the mechanism.
Adaptive placement is the operational definition of TaRL.
Prem's clearest framing of what the Pratham-family methodology has always called Teaching at the Right Level. Imagine a Class 5 student sitting in a Class 5 lesson while reading at Class 2 — she is lost, and the teacher's instruction is irrelevant to her actual cognitive state. The model has to start her at Class 2, and route her up one rung at a time, regardless of which standard she is enrolled in.
On scale, the math is grim. India has roughly three hundred million children in school. If half are lagging — which is what the ASER reading numbers say — that is one hundred and fifty million children sitting in lessons they cannot follow. Adaptive placement, whether done by a teacher in a small-group setting (Pratham Education's original Read India methodology) or by software at thirty-five minutes a day (MindSpark in the Lucknow build), is the only mechanism the conversation describes that scales without adding teachers India does not have.
COVID compressed eight years of smartphone penetration into eight months.
The most underappreciated data point in the conversation. When the lockdown hit in March 2020 and Pratham InfoTech wanted to move its Lucknow programme to remote delivery, the first thing they did was count. Of twelve thousand enrolled students, twenty-two percent of families had a mobile phone, and only twelve to thirteen percent had a smartphone. Eight months later, the smartphone number was at seventy percent. Eight months. Among government-school families in Uttar Pradesh.
The two readings of this are different. The optimistic reading is that the parents — government-school parents, not affluent ones — invested aggressively in the child's education because they understood it mattered, and reached for the only available delivery rail. The pessimistic reading is that the delivery rail collapsed pedagogically anyway because the children dissociated from a year of fragmented online instruction. Prem holds both readings simultaneously. The smartphone exists; the school still has not figured out how to use it.
Bangladesh did not have more money. It had a better answer.
Pankaj's flattest line in the conversation. Asked whether India has caught up on girl-child education, he says no — and explains. Bangladesh is now at per-capita-income parity with India, largely because it solved girl-child education twenty-five years ago. India has spent more per capita on education than Bangladesh for the same period. The result is that India still does not have parity, and Bangladesh does.
The implication for anyone who treats education spending as the policy lever is unwelcome. Allocation is not the binding constraint. Conversion is. Money spent on a model that does not produce results produces no results, regardless of the amount. The Bangladesh story should appear in every Indian primary-education debate as a permanent footnote: spending more was not the missing thing.
The DIB worked. Both said it was too costly to repeat at a hundred times the scale.
The closing exchange is the most honest stretch of the conversation. Vishal asks whether India needs more DIBs. Pankaj says yes in principle — paying on outcome is fundamentally desirable. Then he says no, in practice, at scale. Testing on a hundred-times-larger sample is itself a complex and expensive operation; the negotiation overhead of a structure with this many stakeholders is high; and the cost-per-outcome of running another DIB at one hundred times the size would be a different ball game.
Prem agrees. The instrument worked. The design needs to be simpler and the testing needs to be cheaper before another one can run at the scale India actually needs. Both are careful to thank the British Asian Trust for starting target-based financing in Indian education. Both make it clear that the next instrument must inherit the principle, not the structure.
Lines worth keeping near your desk.
The jargon, unpacked.
Some of these are sector standard; some are specific to this conversation. Skim, mark the unfamiliar, return when needed.
Check what you actually retained.
Try the answer first; click to reveal. The point is to notice where the conversation is fuzzy in your memory.
Five questions worth sitting with.
No correct answers. Type into the boxes — your responses are saved locally in your browser.
Pankaj's design starts from the operator who exists, not the operator he wishes for. Where in your own system are you wishing for a non-existent operator — and what would you have to redesign if you accepted the average one?
"Pleasant, not fun." Which of your engagement metrics is actually measuring entertainment rather than mastery? What would the metric look like if it measured the internal signal "I learnt / I did / I made progress"?
The classroom is three interactions, not one. In your work, what are the three (or four) edges between actors, and which one is undesigned because nobody has named it?
The DIB formalised the monitoring but did not change the pedagogy. When you introduce an accountability layer on top of an existing operation, what is the thing the layer is honestly transforming — and what is the thing you are hoping it will transform but probably won't?
Bangladesh did not have more money than India. It had a better answer. Where is your team confusing allocation efficiency with conversion efficiency — and which problem are you actually trying to solve next quarter?
Where to push back.
The strongest version of each disagreement, written to be persuasive — not to win.
"Government schools at scale will out-perform NGO models on cost-per-outcome."
The counter: a government primary teacher in India costs three-to-five times what a Gyan Shala paraprofessional costs, with weaker measured outcomes against the same ASER comprehension test. The cost-per-outcome arithmetic is therefore reversed: the NGO model is cheaper per literate child, not more expensive. Government schools win on cost-per-enrolled child (which was the access war's metric) but lose on cost-per-outcome (which is the metric that matters now). The harder version of this objection is not that NGOs are too expensive but that the government cannot adopt NGO methods because its labour structure (permanent recruitment, union protections, training pipelines) is set up for the old metric. That is a constraint on adoption, not on the underlying economics.
"200K is impressive but irrelevant. The math problem is 300 million children."
The push: pilots at this scale do not have to scale themselves; they have to demonstrate replicability of the operating model. NITI Aayog has, by Prem's account on tape, adopted the MindSpark intervention and is replicating it elsewhere — that is the path by which a successful 200K pilot becomes a 20-million programme. The harder critique is structural: at 300 million children, the rate-limiting step is not the design of the intervention but the political-economy of state-government adoption, which the DIB's results-based contracting does not address. The pilot proves the engineering; nothing in the pilot proves the politics. That is the gap that the next instrument has to fill.
"The 200K impact is short-lived without curriculum reform."
The push: the durability of FLN gains beyond the intervention window is the dataset that nobody — not the British Asian Trust, not the delivery NGOs, not the testers — has invested in producing. A Class-5 cohort that learnt to read at comprehension level under the DIB and then sat in a Class-6 government classroom that taught at neither the right level nor with structured pedagogy is the cohort whose four-year fade-out should be measured. The DIB design did not include this; it could not, given the four-year clock. The next instrument should treat post-intervention follow-up as a first-class outcome and pay against five-year retention of learning level, not just two-year gain.
"Outcomes-based finance buys outcomes for the funder, not impact for the child."
The push: this is a real risk in weak test designs and is the standard critique of every outcomes-based contract in education. The QEI DIB defended against it by handing the test to a third-party tester with its own framework, by negotiating the test as comprehension rather than literacy, and by using a comparator group of low-cost private schools rather than the delivery partner's own retained children. Those structural protections reduce the gaming surface but do not eliminate it. The harder steelman is that any sufficiently rich measurement instrument can be inverted into a teaching script, and that even comprehension tests can be drilled. The remedy is rotating test designs across the contract window — which the QEI DIB did not do, and which the next instrument should.
"The MindSpark fivefold-lift is a software artefact, not a learning artefact."
The push: this is the cleanest measurement-credibility objection in the conversation, and it deserves more weight than it gets. A platform's own scoring engine flatters a programme that uses the platform; an external comprehension test does not. The conversation conflates the two. The defensible version of the MindSpark result is the third-party test, not the platform's internal scale. The honest reading is that MindSpark's daily-data discipline produced an operationally tight programme, which the third-party comprehension test then validated — but the headline "5x lift" is the platform's number, and the audited number is the comprehension number. Future DIBs should publish both, and lead with the audited one.
Three angles on Monday morning.
If you don't work in education, take the operating logic.
If you're an operator
- Audit your frontline-operator assumption. Write down the specific capability your programme silently requires; benchmark against the actual median capability of the people you have. If the gap is non-trivial, redesign — not retrain.
- Specify the three interactions in your delivery model and check which is undesigned. The undesigned interaction is usually carrying the largest hidden inefficiency.
- Replace engagement metrics with mastery metrics. The behavioural signal "the user did the thing" is weaker than the internal signal "the user got the thing right" — instrument the second.
- If you can instrument at daily resolution, use the data for daily propagation of best-practice from top performers to bottom performers, not for monthly reporting.
- Pre-print the script. Make average performance produce the intended outcome.
If you're a funder
- Be honest with yourself about what an outcomes-based contract transforms. It transforms governance and reporting. It does not transform pedagogy. Price the governance overhead at the delivery partner; do not assume it pays for itself.
- The comparator-group negotiation is the political moment of the contract. If the natural benchmark (a government school, the incumbent provider, the BAU operation) refuses to be measured, that refusal is data — treat it as such.
- Test design upstream of payment is more important than payment design. Pay for the harder test (comprehension over literacy, retention over gain), not the easier one.
- Build five-year retention measurement into the next instrument, not just intervention-window gain. A four-year DIB does not prove a child still reads at the level she learnt to.
- Rotate test instruments across the contract window. Any sufficiently rich measurement can be drilled into; preventing that requires test rotation, not test richness.
If you're a policymaker
- The access war is won (98.5% enrolment). The learning-outcomes war has not started in earnest. Move the funders, the metrics, and the operating models that were built for the previous war out of the way for the new one.
- Adopt adaptive placement (TaRL) as a routing instrument, not a pedagogy. The lesson the child receives does not need to change; the placement upstream of the lesson does.
- Treat teacher capability as a designed-around variable, not a trained-up variable. The shortage of trained teachers will not close in your political timeline; structured pedagogy will.
- Allocation is not the binding constraint; conversion is. Stop debating budget size against Bangladesh and start debating the mechanism that converts budget into outcome.
- For DIB-style instruments to scale tenfold, the testing-and-monitoring overhead has to drop by an order of magnitude. Fund the test-design work, not just the delivery work, in the next instrument.
How we got here.
The whole conversation, searchable.
Click a timestamp to open YouTube at that moment. Click any line to highlight it (yellow). Highlights and notes save in this browser only.
00:00 in the page to seek.