Episode 37 · The UpStream Life · Vishal Krishna in conversation with Pankaj Jain (Gyan Shala) and Prem (Pratham InfoTech Foundation)

Two delivery partners on what an eleven-million-dollar education bond actually asked of the classroom.

The Quality Education India Development Impact Bond ran from 2018 to 2022, paid against measured learning outcomes, and routed eleven million dollars through four NGOs into the hands of two-hundred-thousand primary-school children. Pankaj Jain built Gyan Shala in 1999 around the realisation that an ordinary village teacher cannot be expected to perform like a college professor — and that the school system has to be designed around the teacher who actually exists. Prem co-built Pratham InfoTech in 1998 around a donation of two hundred used 486 computers and the belief that a hybrid of technology and tradition is the only model that survives Indian scale. This conversation is what those two playbooks looked like when paid for on results.

Guests Pankaj Jain (CEO, Gyan Shala) and Prem (CEO, Pratham InfoTech Foundation)· Host Vishal Krishna· Length 62 min· Recorded April 2023 · remote
Video thumbnail
How Gyan Shala and Pratham InfoTech Foundation impacted the lives of 200K school students
Embedding off · opens YouTube

In sixty seconds.

India's school-education problem is no longer access. The ASER number Prem reads off the page is 98.5%: that is enrolment. The companion number — also from ASER — is that fewer than half the children in Class 5 in a government school can read a Class 2 text. Enrolment is solved. Reading is not. The whole conversation sits inside that gap.

Pankaj Jain's answer is structural. Gyan Shala does not try to upskill the teacher into a college professor. It accepts the village 12th-pass woman as the teacher who actually exists, and engineers the classroom — three kinds of interaction, structured pedagogy, written-friendly furniture, three to four hours, never more — around her capacity. The yield against a low-cost private-school comparator group ran two to three times above the contracted target across four years.

Prem's answer is hybrid. Two hundred used 486 computers in 1998 became, by 2019, fifty-five Lucknow government schools equipped with twelve-hundred Chromebooks running MindSpark — an adaptive-learning platform built with Educational Initiative. The Class 5 child reading at Class 2 logs in at Class 2; she does not stay lost in a Class 5 lesson. The treatment group rose from a nine-scale-point baseline to forty-three to forty-five. A five-fold lift, audited by a third party, was the bond's currency. Neither delivery partner thought the DIB redesigned their pedagogy. Both said it formalised their monitoring and forced the testing to be done by someone they did not pick.

Where to land in the conversation.

Each chapter opens YouTube at that timestamp in a new tab.

Six ideas to carry into your own work.

The structural arguments inside the conversation, lifted out so they travel. Each one is the kind of thing you can quote in a programme review on Tuesday.

01

Design for the teacher you actually have.

Pankaj's plainest line. Most school reform programmes assume the teacher will rise to meet the curriculum. Gyan Shala starts from the opposite premise: a 12th-pass village woman with limited training is the teacher who shows up, and the entire classroom — script, materials, timing, the three kinds of interaction — has to be engineered around her capability. Structured pedagogy is the operational name for that decision.

If your programme requires an above-average operator, your programme has not been designed yet. It has been wished into the population it does not have.
02

Pleasant, not fun.

The classroom that retains a child is not the one that entertains her. It is the one in which the pleasure of empowerment — the unmistakeable feeling of having learnt something — happens reliably. Fun is a category mistake; it makes "fun" the objective and learning the by-product. Pankaj reverses the dependency. The class must lead the child to the experience of learning; from that experience the willingness to come back follows.

When a programme starts confusing engagement with outcome, the metric drifts. Pleasure-of-mastery is the engagement that compounds.
03

The three classroom interactions.

Inside any class period, exactly three interactions happen: child with teacher, child with other children, child with the learning material. Most reform programmes obsess about the first, occasionally upgrade the third, and ignore the second. Gyan Shala designs all three on purpose, in the order they affect outcome. The classroom is not a relationship — it is a network with three edges, each of which has to be loaded.

The interaction model is the curriculum. When you change the design of any one edge, you have changed the curriculum, whether the lesson plan says so or not.
04

Teaching at the right level.

Prem's adaptive-learning answer to the ASER gap. A child in Class 5 who reads at Class 2 will not learn from a Class 5 lesson; she will dissociate. MindSpark routes her into a Class 2 path, and lifts her one rung at a time. The diagnostic and the placement are the program. The lesson is just the consequence. It is the only mechanism the conversation describes that works at scale without adding teachers.

Adaptive routing is a planning instrument, not a learning instrument. The lesson the child gets does not need to be smart — the placement upstream of the lesson does.
05

Outcomes-based finance, viewed from the delivery side.

The DIB did not redesign Gyan Shala's pedagogy or Pratham InfoTech's hybrid model. What it did, in both their words, was three things: (i) lock the contractual learning outcomes to the state curriculum's expected outputs, (ii) formalise process monitoring that was previously informal, and (iii) hand the testing to a third party with its own framework. The first re-phased the curriculum slightly. The second cost time. The third required negotiation. None of them changed what worked.

A funding instrument changes the way work is governed, not the way it is done. The mistake is to think payment-on-outcomes will redesign the model — it formalises it.
06

Money is not the constraint. Result is.

Pankaj's flat correction of the Bangladesh comparison. India has spent more per capita on education than Bangladesh for decades, and Bangladesh moved its girl-child cohort to per-capita parity twenty-five years ago. The constraint was never the budget; it was the inability to convert a budget into an outcome. Outcomes-based finance is, in that frame, a structural answer to the right problem.

Allocation efficiency and conversion efficiency are different problems. Most policy debates conflate them. Bangladesh's lesson is that the second is the binding one.

Fifteen things to actually walk away with.

Each one carries the timestamps where the moment lives, and a transferable note for work that isn't education.

01

The teacher capability question is the design question.

Pankaj's argument is sharper than the standard NGO line about "the teacher problem". He says explicitly that mainstream education thinkers have spent twenty years opposing structured pedagogy on the grounds that it de-skills the teacher. His response: the average primary-school teacher in rural Gujarat is not being de-skilled; she is being respected. The system being asked of her — to behave like a four-year-trained college professor managing a forty-child mixed-ability classroom — is the de-skilling, because it is impossible.

The Gyan Shala design starts from the actual skill level and works backwards. Lesson scripts are pre-structured. Materials are pre-printed. The three interactions are pre-engineered. The teacher does not need to be the curriculum designer; she needs to be a competent executor of a curriculum designed by someone who understands how a child's mind acquires language, deductive thought, and inductive thought.

Beyond education. Any system that depends on its frontline workers being above average is a system that has not been designed. The designer's job is to make average performance produce the intended outcome.
02

Three to four hours of school is enough, if the three to four hours are right.

The unusual claim. Pankaj says a child can achieve mastery of grade-level learning in three to four hours a day, in a classroom that is non-threatening and non-paralysing, with the three interactions running properly. Gyan Shala runs in that frame: a focused supplementary or parallel session that the slum child can fit alongside whatever else her day demands.

The implication for policy is uncomfortable. Indian government schools run six to seven hours; private schools run longer. If three to four well-designed hours match or exceed the outcomes of six to seven undesigned hours, the question is not how much instruction time exists but how it is being used. The conversation does not labour the point. It does not have to.

Beyond education. Time-on-task is not a proxy for output. A shorter, well-designed cycle frequently beats a longer, under-designed one. Audit the cycle before you ask for more time.
03

Drop-out is mostly a value problem, not a poverty problem.

Vishal's framing question was the standard one: poor families pull children out of school because they need them to earn. Pankaj reframes it. In his read of Gyan Shala's experience, ninety percent of dropouts are not earning-pressure — they are children who concluded the school is not giving them enough to stay. The five-to-ten percent who drop out for economic or social reasons are real, but they are the minority case.

The implication is that drop-out is downstream of learning outcomes, not upstream of them. Pay families to keep children in unhelpful schools and the children still leave; build a school in which children visibly learn and the families do not need paying. The lever is the classroom, not the cash transfer.

Beyond education. When the visible reason for an attrition pattern (poverty, distance, social pressure) is structural, look for the invisible reason (the experience itself produces no return) before designing the next intervention.
04

"Pleasant," and the careful avoidance of "fun."

The most precise small moment in the conversation. Vishal asks how to make school fun; Pankaj declines the word. Fun risks becoming the objective, and an objective of fun is satisfied by sweets and games, not by learning. Pleasant is the right frame, and the pleasure that retains children is specifically the pleasure of mastery — the unmistakable internal signal that you have learnt something.

The distinction is not pedantic. Twenty years of edtech in India has been built on the fun frame and largely failed. Prem says explicitly later in the conversation that the assessment-and-grading edtech company he saw fail "very weird man" was operating in that frame. Pleasure-of-mastery is what carries a child back into the room the next day.

Beyond education. Engagement metrics are not retention metrics. The signal that compounds is the user's internal experience of "I learnt / I did / I made progress," not the user's expressed enjoyment.
05

The classroom is a three-edge network, not a relationship.

This is the most operational moment in Pankaj's section. Inside a class period, exactly three interactions happen: child with teacher, child with other children, child with the learning material. He calls them out as a designable triple and notes that all three have to be engineered in coordination, not in isolation.

Most reform programmes obsess about the first edge (teacher-training initiatives), occasionally touch the third (better textbooks), and almost never touch the second (peer interaction). Gyan Shala's classroom design is unusual in that it loads all three on purpose, with the second edge contributing a non-trivial share of the outcome. The Class 5 child does not learn only from the teacher; she learns from the Class 5 child next to her, in a structured peer arrangement that the curriculum designer has set up.

Beyond education. Whenever your system has three or more interacting actors, audit which interactions you have actually designed. The undesigned ones usually carry the largest unrealised gain.
06

Mind, not skill: how learning is actually modelled.

Pankaj pushes back when Vishal compresses his model into "deductive learning". His correction is patient. Language is its own thing — children acquire language; they do not learn it the way they learn mathematics. Mathematics is most cleanly the home of deductive thinking. Science is the home of inductive thinking. The classroom designer has to understand these three different ways the human mind engages with knowledge, and design class processes accordingly.

The seemingly academic distinction is structural. A class that teaches mathematics with the cadence of a language class will fail; a class that teaches a language with the cadence of a science class will fail differently. Gyan Shala's curriculum is built on the recognition that these are three different operations on the mind, and that an effective classroom switches between them deliberately.

Beyond education. "Learning" is not one process. Acquisition, deduction, induction are different cognitive operations, and design that treats them identically delivers diminishing returns on each.
07

ASER 98.5%: the access war is won. The learning war has not started.

Prem's number is exact. The Annual Status of Education Report — Pratham Education Foundation's flagship survey, running for fifteen years at the time of recording — puts primary-school enrolment at 98.5%. India has solved the access problem that consumed two decades of policy attention. The companion number, also from ASER, is that fewer than half of Class 5 children in a government school can read a Class 2 text. Five years of schooling produces a child who is, in literacy terms, three years behind.

The diagnostic implication is that primary education in India is no longer an enrolment problem and is now a learning-outcomes problem. The two require different instruments. Enrolment was moved by mid-day meals, school construction, RTE-style entitlements. Learning will be moved, if at all, by classroom design, adaptive placement, and outcomes-based monitoring. The DIB is a first instrument calibrated for the second war.

Beyond education. When you cross a threshold (98.5%), the constraint moves. The funders, the metrics, and the operating model that solved the previous war are not the ones that win the new one.
08

The DIB is a financing instrument, not a pedagogy.

Both Pankaj and Prem refuse the leading question. Vishal opens by asking how the Quality Education India DIB changed the way they teach. Pankaj corrects: a bond is a financing mechanism. It has nothing to do with the education per se. What worked in a classroom worked whether it was funded by a DIB, a grant, or a government cheque; the design of the learning process did not change because the funding structure did.

What did change was governance. The DIB locked the contractual learning outcomes to a negotiated subset of the state curriculum, formalised process monitoring that had previously been informal, and required testing to be done by a third party with its own framework. None of this is pedagogy. All of it is overhead, sometimes useful, sometimes costly. The intellectual honesty here is rare — outcomes-based finance is often sold as a pedagogy transformer, and it is not.

Beyond education. A funding mechanism transforms governance and reporting. It rarely transforms operations. Buyers of new financing structures should price the governance overhead, not the imagined operational lift.
09

The comparator group revealed who held the system to account.

An overlooked detail. The DIB structure required Gyan Shala's progress to be benchmarked against a comparison group of similar socioeconomic profile. The ideal comparator would have been a government school. The government — Pankaj says with diplomatic understatement — would not agree to be the benchmark. The comparator that was negotiated was a set of low-cost private schools serving children of broadly similar background.

This is two findings, not one. First, the government's reluctance is itself evidence of where accountability for outcomes sits in the Indian primary-education system: nowhere stable. Second, the comparator that was chosen is the harder one — low-cost private schools are paid for by aspirational parents and have stronger incentives than government schools to produce visible literacy. Gyan Shala outperforming the comparator by two-to-three-times the contracted target is therefore the stronger version of the result, not the weaker one.

Beyond education. A comparator-group choice is a political act, not a methodological one. Who is willing to be benchmarked tells you who is confident about their result and who is not.
10

The test became a test of comprehension, not literacy.

A second overlooked detail of the bond's design. The independent tester chosen by the British Asian Trust did not run a literacy test (read these letters, decode these sounds). It ran a comprehension test (read this passage, understand it, answer questions on it). Pankaj is clear that this is a stronger test, and that the Class 5 children failing the ASER literacy test would have failed this one more severely.

The choice mattered. A test of comprehension forces the curriculum to teach for comprehension; a test of literacy permits the curriculum to teach for decoding. Gyan Shala's outperformance on the comprehension test is a stronger claim than the same outperformance on a literacy test would have been, because comprehension is harder to game with surface drilling. The DIB structure, by handing the testing to a third party with sophisticated instruments, raised the bar for the delivery partner and made the result more credible.

Beyond education. The test design upstream of payment is more important than the payment itself. Outcomes-based contracts where the test is weak deliver gaming; outcomes-based contracts where the test is hard deliver real outcomes.
11

Lucknow, Mission 35, and the daily-data discipline.

Pratham InfoTech's Lucknow build is the most operationally specific stretch of the conversation. Fifty-five Uttar Pradesh government schools, twelve hundred Chromebooks, a learning lab per school with twenty-to-twenty-five computers in one classroom, MindSpark adaptive software underneath. Fifty-five further schools chosen as a matched control. The treatment was a daily target — thirty-five minutes of MindSpark per child per day. They called it "Mission 35." A schedule that took weeks of negotiation with principals and management to fit into the school day.

The daily-data discipline was the hidden lever. Because MindSpark logged every interaction, Pratham could rank schools by daily performance, identify a top five and a bottom five, and propagate the practices of the top five to the rest on the same day's cycle. That feedback loop was tighter than anything paper-based monitoring could have produced. The treatment group ended up at a forty-three-to-forty-five-scale-point baseline versus the control at nine. The fivefold lift is the headline; the daily ranking-and-propagation is the mechanism.

Beyond education. When you can instrument a programme at daily resolution, the right action is not better strategy but faster feedback. Tighten the loop before you change the model.
12

Adaptive placement is the operational definition of TaRL.

Prem's clearest framing of what the Pratham-family methodology has always called Teaching at the Right Level. Imagine a Class 5 student sitting in a Class 5 lesson while reading at Class 2 — she is lost, and the teacher's instruction is irrelevant to her actual cognitive state. The model has to start her at Class 2, and route her up one rung at a time, regardless of which standard she is enrolled in.

On scale, the math is grim. India has roughly three hundred million children in school. If half are lagging — which is what the ASER reading numbers say — that is one hundred and fifty million children sitting in lessons they cannot follow. Adaptive placement, whether done by a teacher in a small-group setting (Pratham Education's original Read India methodology) or by software at thirty-five minutes a day (MindSpark in the Lucknow build), is the only mechanism the conversation describes that scales without adding teachers India does not have.

Beyond education. When the population is bigger than the trained-operator supply, the only scalable model is one that diagnoses and routes the individual without requiring an above-average operator at each interaction.
13

COVID compressed eight years of smartphone penetration into eight months.

The most underappreciated data point in the conversation. When the lockdown hit in March 2020 and Pratham InfoTech wanted to move its Lucknow programme to remote delivery, the first thing they did was count. Of twelve thousand enrolled students, twenty-two percent of families had a mobile phone, and only twelve to thirteen percent had a smartphone. Eight months later, the smartphone number was at seventy percent. Eight months. Among government-school families in Uttar Pradesh.

The two readings of this are different. The optimistic reading is that the parents — government-school parents, not affluent ones — invested aggressively in the child's education because they understood it mattered, and reached for the only available delivery rail. The pessimistic reading is that the delivery rail collapsed pedagogically anyway because the children dissociated from a year of fragmented online instruction. Prem holds both readings simultaneously. The smartphone exists; the school still has not figured out how to use it.

Beyond education. A capacity that did not exist twelve months ago can be in the hands of seventy percent of your target users today. Programme design that assumes hardware availability as a multi-year roadmap variable is over-cautious in a way that costs.
14

Bangladesh did not have more money. It had a better answer.

Pankaj's flattest line in the conversation. Asked whether India has caught up on girl-child education, he says no — and explains. Bangladesh is now at per-capita-income parity with India, largely because it solved girl-child education twenty-five years ago. India has spent more per capita on education than Bangladesh for the same period. The result is that India still does not have parity, and Bangladesh does.

The implication for anyone who treats education spending as the policy lever is unwelcome. Allocation is not the binding constraint. Conversion is. Money spent on a model that does not produce results produces no results, regardless of the amount. The Bangladesh story should appear in every Indian primary-education debate as a permanent footnote: spending more was not the missing thing.

Beyond education. When a peer has overtaken you on less budget, the budget is not the issue. The conversion mechanism is. Spend the next debate on the mechanism, not the allocation.
15

The DIB worked. Both said it was too costly to repeat at a hundred times the scale.

The closing exchange is the most honest stretch of the conversation. Vishal asks whether India needs more DIBs. Pankaj says yes in principle — paying on outcome is fundamentally desirable. Then he says no, in practice, at scale. Testing on a hundred-times-larger sample is itself a complex and expensive operation; the negotiation overhead of a structure with this many stakeholders is high; and the cost-per-outcome of running another DIB at one hundred times the size would be a different ball game.

Prem agrees. The instrument worked. The design needs to be simpler and the testing needs to be cheaper before another one can run at the scale India actually needs. Both are careful to thank the British Asian Trust for starting target-based financing in Indian education. Both make it clear that the next instrument must inherit the principle, not the structure.

Beyond education. A pilot instrument that worked at five organisations and 200K beneficiaries is not the instrument that will work at fifty organisations and twenty million. The principle scales; the structure does not. Re-engineer the structure before you re-issue.

Lines worth keeping near your desk.

Our expectation from the school teacher is what you'd hold from the four professors of a higher education college. This is not something which can be doable. Pankaj Jain · 05:14
Nothing is more pleasant than learning. The feeling of empowerment which comes out of learning is what retains the child. Fun is the wrong word. Pankaj Jain · 14:49
Forty-four percent of children in government school in Class 5 cannot read a Class 2 text. They have spent five years of their schooling time and they are not able to read. Prem · 19:49
A bond is essentially a financing mechanism. It has nothing to do with the education per se. The design of the learning process did not change for us just because it is a development impact bond. Pankaj Jain · 23:23
Money is not enough. You have to succeed in using the money to get the outcome that you want. Our problem is not that we are not spending money. Our problem is that we are not getting the result. Pankaj Jain · 53:11

The jargon, unpacked.

Some of these are sector standard; some are specific to this conversation. Skim, mark the unfamiliar, return when needed.

DIB
acronym
Development Impact Bond. A results-based financing structure in which an outcomes funder repays a risk investor only if independently-measured outcomes are achieved by the delivery partner. The QEI DIB ran 2018-2022, eleven million dollars, four delivery NGOs, ~200K children.
QEI DIB
programme
Quality Education India Development Impact Bond. Launched 2018 by the British Asian Trust as the world's largest education DIB at the time. The instrument under which both Gyan Shala and Pratham InfoTech operated for four years.
British Asian Trust
funder
UK-headquartered charity, founded 2007. The outcomes funder (and convener) for the QEI DIB. Companion to Episode 36, which sits in the funder seat to this episode's delivery seat.
ASER
acronym
Annual Status of Education Report. Pratham's flagship household survey of Indian primary-school learning levels, running since 2005. The source of the 98.5% enrolment number and the <50% Class-5-reading-Class-2 number Prem quotes.
FLN
acronym
Foundational Literacy and Numeracy. The category of outcomes — reading-with-comprehension and basic arithmetic by end-of-Class-3 — that NEP 2020 made an explicit national mission. The QEI DIB's testing instrument measured comprehension, not just decoding.
TaRL
methodology
Teaching at the Right Level. Pratham-family pedagogy in which children are assessed and grouped by current ability rather than enrolled grade, then taught at their level until they progress. The conceptual ancestor of MindSpark's adaptive routing.
Structured pedagogy
approach
A school design approach in which lesson scripts, materials, assessments, and teacher routines are pre-engineered. Pankaj defends it against accusations of "de-skilling the teacher" by re-framing: it respects the teacher's actual capability rather than wishing for a non-existent one.
Paraprofessional teacher
role
A 12th-pass woman from the local community, hired and trained as a classroom teacher, at a fraction of the cost of a state-trained graduate. The unit operator of Gyan Shala's slum-school model.
MindSpark
platform
Adaptive-learning software built by Educational Initiative (Ahmedabad). Diagnoses each child's level on login, routes her into appropriate exercises, and propagates progress. The technology layer for Pratham InfoTech's Lucknow build.
Educational Initiative
organisation
Ahmedabad-based education-assessment-and-content company. Built MindSpark. Pratham InfoTech's third partner in the Lucknow DIB intervention, alongside the UP state government.
Mission 35
programme name
Pratham InfoTech's internal target: every child in a treatment school logs thirty-five minutes per day on MindSpark, against an annual calendar of about twenty hours per term. The operational lever that ran the Lucknow programme.
NEP 2020
policy
India's National Education Policy 2020. Reframes primary-school instruction around foundational literacy and numeracy, multilingual instruction in the early grades, and outcomes-based assessment. The policy context against which both delivery partners operate.
Comparator group
methodology
The non-treated population whose learning trajectory is benchmarked against the treated population. In the QEI DIB, the government refused to serve as comparator; the negotiated comparator was a set of low-cost private schools.
Supplementary class
model
A delivery model in which the NGO runs an additional class outside or alongside the government school, rather than replacing it. The default Gyan Shala model in the slums Pankaj describes.
Acquired vs learnt language
distinction
Pankaj's correction to a question about deductive learning. A child acquires language through interaction; she learns mathematics through deductive thought and science through inductive thought. Three different cognitive operations; the curriculum has to switch between them.
Ayog
institution
Short for NITI Aayog — the Government of India's policy think-tank since 2015. Prem mentions that NITI Aayog has adopted and is replicating the MindSpark model in other parts of India following the Lucknow result.

Check what you actually retained.

Try the answer first; click to reveal. The point is to notice where the conversation is fuzzy in your memory.

Q1
What is the precise ASER number for primary-school enrolment that Prem quotes, and what is the companion learning-outcome number?
Enrolment is 98.5%; only 1.5% of children are out of school. The companion number is that 44% of Class 5 children in government schools can read a Class 2-level text — meaning more than half cannot. Enrolment is solved; reading is not.
Q2
Why does Pankaj reject the standard objection that "structured pedagogy de-skills the teacher"?
Because the alternative implicitly expects the village 12th-pass paraprofessional teacher to perform like a four-year-trained college professor managing forty mixed-ability children. That is the de-skilling — it is impossible. Structured pedagogy starts from the teacher's actual capability and engineers around it. It respects the operator who exists rather than wishing for the one who does not.
Q3
Why does Pankaj refuse the word "fun" and prefer "pleasant"?
Because fun risks becoming the objective of the classroom — and "fun" is satisfied by sweets and games, not by learning. The pleasure that actually retains a child is the pleasure of mastery: the internal signal "I learnt something today." Pleasant is the correct framing because it points at that signal rather than at entertainment.
Q4
What are the three classroom interactions Pankaj insists must be designed in coordination?
Child with teacher, child with other children, child with the learning material. Most reform programmes obsess about the first edge (teacher training) and occasionally touch the third (better textbooks); the second edge (peer interaction) is almost always left undesigned. Gyan Shala loads all three on purpose, with peer interaction contributing non-trivially.
Q5
What two changes did the DIB impose on Gyan Shala and Pratham InfoTech, by their own description, without changing pedagogy?
First, the contractual learning outcomes were locked to a negotiated subset of the state curriculum, which slightly re-phased the curriculum over time. Second, process monitoring that had been informal had to be formalised — costly in teacher and supervisor time. Third (a third, not a "change"), the testing was handed to an independent third party with its own framework, requiring time-consuming negotiation. None of these changed the design of teaching.
Q6
Why was the comparator group in Gyan Shala's measurement a set of low-cost private schools rather than government schools?
Because the government refused to allow government schools to be benchmarked against an NGO programme. The DIB then negotiated low-cost private schools serving similar-background children as the comparator. This is, if anything, the harder benchmark — those schools are paid for by aspirational parents and have incentives to deliver visible literacy. Gyan Shala's 200-300% outperformance against this comparator is the stronger version of the result.
Q7
What is the operational scale of Pratham InfoTech's Lucknow DIB intervention?
Fifty-five UP government schools as the treatment group, fifty-five matched schools as control, twelve hundred Chromebooks across the treatment group, twenty-to-twenty-five computers per learning lab, MindSpark adaptive software, and a target of thirty-five minutes of usage per child per day — branded internally "Mission 35." The treatment group rose from a 9-scale-point baseline to 43-45 scale points; the control stayed flat. A roughly 5x lift.
Q8
What is Teaching at the Right Level, in one sentence, and why does it scale?
Assess each child's actual reading and arithmetic level on entry; route her into instruction at that level regardless of her enrolled grade; advance her one rung at a time. It scales because it does not require an above-average teacher at every interaction — adaptive software (or grouped small-group instruction) can do the routing, and the teacher's role compresses to facilitation rather than diagnosis.
Q9
What COVID-era data point captures the speed of smartphone penetration in government-school families in Lucknow?
Pratham InfoTech surveyed its twelve thousand enrolled students at lockdown. Twenty-two percent of families had any mobile phone; twelve-to-thirteen percent had a smartphone. Eight months later the smartphone number had risen to seventy percent. The hardware floor moved in eight months, not eight years.
Q10
What does Pankaj's Bangladesh comparison establish about the Indian education-policy debate?
India has spent more per capita on education than Bangladesh for decades. Bangladesh moved its girl-child cohort to per-capita-income parity twenty-five years ago. India has not. The binding constraint is therefore not allocation but conversion — the ability to translate a budget into a measurable outcome. Policy debates that conflate the two solve the wrong problem.
Q11
Both Pankaj and Prem support DIB-style instruments in principle but flag a specific constraint at scale. What is it?
Cost and complexity of the testing-and-monitoring overhead. A QEI-sized DIB needed independent testers, formalised process documentation, comparator-group selection, multi-stakeholder negotiation — all of which run linearly or worse with scale. To do a hundred-times-larger DIB, the structure has to be simpler and the testing has to be cheaper. The principle scales; the current structure does not.
Q12
What is the difference between the QEI DIB's reading test and an ASER literacy test?
ASER tests literacy as decoding — can the child convert text into the right sounds. The QEI DIB's independent tester ran a comprehension test — can the child read a passage, understand it, and answer questions on it. Comprehension is harder, less gameable with surface drilling, and produces a stronger result when the delivery partner outperforms. The test design upstream of payment matters more than the payment itself.

Five questions worth sitting with.

No correct answers. Type into the boxes — your responses are saved locally in your browser.

Pankaj's design starts from the operator who exists, not the operator he wishes for. Where in your own system are you wishing for a non-existent operator — and what would you have to redesign if you accepted the average one?

"Pleasant, not fun." Which of your engagement metrics is actually measuring entertainment rather than mastery? What would the metric look like if it measured the internal signal "I learnt / I did / I made progress"?

The classroom is three interactions, not one. In your work, what are the three (or four) edges between actors, and which one is undesigned because nobody has named it?

The DIB formalised the monitoring but did not change the pedagogy. When you introduce an accountability layer on top of an existing operation, what is the thing the layer is honestly transforming — and what is the thing you are hoping it will transform but probably won't?

Bangladesh did not have more money than India. It had a better answer. Where is your team confusing allocation efficiency with conversion efficiency — and which problem are you actually trying to solve next quarter?

Where to push back.

The strongest version of each disagreement, written to be persuasive — not to win.

"Government schools at scale will out-perform NGO models on cost-per-outcome."

Pankaj's and Prem's gains are real, but the cost base of an NGO programme — paraprofessional teachers paid from philanthropic funds, third-party testing, custom curriculum — is a luxury the government-school system cannot afford at national scale.

The counter: a government primary teacher in India costs three-to-five times what a Gyan Shala paraprofessional costs, with weaker measured outcomes against the same ASER comprehension test. The cost-per-outcome arithmetic is therefore reversed: the NGO model is cheaper per literate child, not more expensive. Government schools win on cost-per-enrolled child (which was the access war's metric) but lose on cost-per-outcome (which is the metric that matters now). The harder version of this objection is not that NGOs are too expensive but that the government cannot adopt NGO methods because its labour structure (permanent recruitment, union protections, training pipelines) is set up for the old metric. That is a constraint on adoption, not on the underlying economics.

"200K is impressive but irrelevant. The math problem is 300 million children."

The QEI DIB touched two-hundred-thousand children over four years — one-fifteen-hundredth of the population of Indian primary-school children. A pilot at this scale cannot speak to national reform.

The push: pilots at this scale do not have to scale themselves; they have to demonstrate replicability of the operating model. NITI Aayog has, by Prem's account on tape, adopted the MindSpark intervention and is replicating it elsewhere — that is the path by which a successful 200K pilot becomes a 20-million programme. The harder critique is structural: at 300 million children, the rate-limiting step is not the design of the intervention but the political-economy of state-government adoption, which the DIB's results-based contracting does not address. The pilot proves the engineering; nothing in the pilot proves the politics. That is the gap that the next instrument has to fill.

"The 200K impact is short-lived without curriculum reform."

The DIB produced measurable gains over four years, but a child who graduated the programme then re-entered an unreformed government secondary school. The long-tail learning trajectory is unfunded and unmeasured.

The push: the durability of FLN gains beyond the intervention window is the dataset that nobody — not the British Asian Trust, not the delivery NGOs, not the testers — has invested in producing. A Class-5 cohort that learnt to read at comprehension level under the DIB and then sat in a Class-6 government classroom that taught at neither the right level nor with structured pedagogy is the cohort whose four-year fade-out should be measured. The DIB design did not include this; it could not, given the four-year clock. The next instrument should treat post-intervention follow-up as a first-class outcome and pay against five-year retention of learning level, not just two-year gain.

"Outcomes-based finance buys outcomes for the funder, not impact for the child."

A DIB pays on measured outcomes. The delivery partner is therefore incentivised to optimise the measurement — to teach to the test, to select children who will move, to manage attrition strategically.

The push: this is a real risk in weak test designs and is the standard critique of every outcomes-based contract in education. The QEI DIB defended against it by handing the test to a third-party tester with its own framework, by negotiating the test as comprehension rather than literacy, and by using a comparator group of low-cost private schools rather than the delivery partner's own retained children. Those structural protections reduce the gaming surface but do not eliminate it. The harder steelman is that any sufficiently rich measurement instrument can be inverted into a teaching script, and that even comprehension tests can be drilled. The remedy is rotating test designs across the contract window — which the QEI DIB did not do, and which the next instrument should.

"The MindSpark fivefold-lift is a software artefact, not a learning artefact."

A treatment group scoring 43-45 against a control at 9 on a scaled MindSpark assessment is impressive — but the scale is calibrated by Educational Initiative, which built MindSpark. The third-party tester evaluated comprehension; the fivefold lift was measured by the platform.

The push: this is the cleanest measurement-credibility objection in the conversation, and it deserves more weight than it gets. A platform's own scoring engine flatters a programme that uses the platform; an external comprehension test does not. The conversation conflates the two. The defensible version of the MindSpark result is the third-party test, not the platform's internal scale. The honest reading is that MindSpark's daily-data discipline produced an operationally tight programme, which the third-party comprehension test then validated — but the headline "5x lift" is the platform's number, and the audited number is the comprehension number. Future DIBs should publish both, and lead with the audited one.

Three angles on Monday morning.

If you don't work in education, take the operating logic.

O

If you're an operator

  • Audit your frontline-operator assumption. Write down the specific capability your programme silently requires; benchmark against the actual median capability of the people you have. If the gap is non-trivial, redesign — not retrain.
  • Specify the three interactions in your delivery model and check which is undesigned. The undesigned interaction is usually carrying the largest hidden inefficiency.
  • Replace engagement metrics with mastery metrics. The behavioural signal "the user did the thing" is weaker than the internal signal "the user got the thing right" — instrument the second.
  • If you can instrument at daily resolution, use the data for daily propagation of best-practice from top performers to bottom performers, not for monthly reporting.
  • Pre-print the script. Make average performance produce the intended outcome.
F

If you're a funder

  • Be honest with yourself about what an outcomes-based contract transforms. It transforms governance and reporting. It does not transform pedagogy. Price the governance overhead at the delivery partner; do not assume it pays for itself.
  • The comparator-group negotiation is the political moment of the contract. If the natural benchmark (a government school, the incumbent provider, the BAU operation) refuses to be measured, that refusal is data — treat it as such.
  • Test design upstream of payment is more important than payment design. Pay for the harder test (comprehension over literacy, retention over gain), not the easier one.
  • Build five-year retention measurement into the next instrument, not just intervention-window gain. A four-year DIB does not prove a child still reads at the level she learnt to.
  • Rotate test instruments across the contract window. Any sufficiently rich measurement can be drilled into; preventing that requires test rotation, not test richness.
P

If you're a policymaker

  • The access war is won (98.5% enrolment). The learning-outcomes war has not started in earnest. Move the funders, the metrics, and the operating models that were built for the previous war out of the way for the new one.
  • Adopt adaptive placement (TaRL) as a routing instrument, not a pedagogy. The lesson the child receives does not need to change; the placement upstream of the lesson does.
  • Treat teacher capability as a designed-around variable, not a trained-up variable. The shortage of trained teachers will not close in your political timeline; structured pedagogy will.
  • Allocation is not the binding constraint; conversion is. Stop debating budget size against Bangladesh and start debating the mechanism that converts budget into outcome.
  • For DIB-style instruments to scale tenfold, the testing-and-monitoring overhead has to drop by an order of magnitude. Fund the test-design work, not just the delivery work, in the next instrument.

How we got here.

1994Pratham Education Foundation founded. "Every child in school and learning well" — the original mission. Mumbai municipal-school slums as the initial geography.
1998Two hundred used 486 computers. ICICI's Mr Vaghul and Mr Kamath donate decommissioned hardware. Pratham takes up the challenge of using them in fifteen Mumbai Municipal Corporation schools. The seed of Pratham InfoTech.
1999Gyan Shala founded. Pankaj Jain starts the Ahmedabad operation around the question Bangladesh forced him to ask: why has Bangladesh moved ahead of India on girl-child education from poor families. Structured pedagogy as the answer.
2004Pratham InfoTech Foundation becomes a separate entity. The education-technology project spins out of Pratham Education with its own board and operations.
2005ASER launches. Pratham's annual household survey of Indian primary-school learning levels begins. The dataset that, fifteen years later, lets Prem read off the 98.5% enrolment and the <50% Class-5-reading-Class-2 numbers from the same page.
2018Quality Education India DIB launches. The British Asian Trust convenes the world's largest education impact bond at the time. Eleven million dollars. Four delivery NGOs. A four-year clock against measured learning outcomes.
2019Lucknow. Pratham InfoTech joins the DIB in year two. Fifty-five UP government schools as treatment, fifty-five as control, twelve hundred Chromebooks, MindSpark, Mission 35.
Mar 2020COVID. Schools close. Pratham InfoTech surveys its twelve thousand enrolled families: 22% have mobile phones, 12-13% have smartphones.
Nov 2020The hardware floor moves. Eight months later, smartphone penetration among Pratham InfoTech's UP government-school families has risen to seventy percent. The remote-delivery rail exists; the pedagogical model for using it does not yet.
2020NEP 2020. India's National Education Policy 2020 frames foundational literacy and numeracy as a national mission. The policy context catches up with what the DIB was already trying to fund.
2022QEI DIB concludes. Four years of measured outcomes. Gyan Shala's treatment cohort: 2-3x the contracted target against the low-cost-private-school comparator. Pratham InfoTech's Lucknow cohort: 9 to 43-45 scale points. NITI Aayog begins replicating the model.
Apr 2023This recording. Pankaj Jain and Prem reflect on what the bond asked of them, what it did not change about their pedagogy, and what the next instrument has to be cheaper and simpler to scale.

The whole conversation, searchable.

Click a timestamp to open YouTube at that moment. Click any line to highlight it (yellow). Highlights and notes save in this browser only.

Shortcuts: / focus search · j/k previous/next segment · h toggle highlight on active line · click any 00:00 in the page to seek.
About this transcript. Captions were pulled from YouTube's auto-generated subtitles and grouped into roughly twelve-second blocks, then cleaned for obvious mis-hearings. The auto-captions consistently render "Gyan Shala" as Gyan Chala/ganshala/gansara, "Pratham InfoTech" as pratham infotech, "Pankaj Jain" as pankaj Jane, "Pankaj-ji" as pankaja, "Prem" as Premier/praining/Prime, "ASER" as user/assar, "MindSpark" as mindspark/Mind spark, "edtech" as attack, "deductive" as directive/Detective, "Quality Education India DIB" as the spelled-out long form, and "bridging the digital divide" as Bridging the digital device. Common fixes were applied; treat as a working transcript, not a verbatim record.

Built as a personal listening tool. Video stays on YouTube.
/listening-lab · ep 37