Principles

Ten Positions on AI in School


Ten positions on what AI should and should not do to learning. Each is stated so that a competent person could reject it, and each carries a demand a head teacher could make of any supplier, ourselves included.

Positions
10
Addressed to
Schools, and anyone selling to them
Status
Open to argument
Cost
None
Index

The teacher’s judgement is final, and disagreement must cost her nothing

In the loop is a phrase most suppliers use, and what it usually amounts to is that the teacher is shown a conclusion and given a button. Authority is not a button. It is the answer to three questions a school can put to any supplier: what does it cost a teacher to disagree, what is her disagreement then treated as, and what becomes of her decision afterwards.

The first is friction. If accepting a recommendation takes one tap and rejecting it requires a typed justification, the design has put a thumb on the scale, and over a term of fifty-five-student sections we would expect the thumb to win. What that produces is not agreement but deference by fatigue, and in every record the school keeps the two look identical. The second is interpretation: a teacher’s disagreement is not a compliance failure, and a school is entitled to ask where it goes and who is shown it. The third is persistence: a decision the system quietly reverses later was never the teacher’s to make.

The objection we take most seriously is that teacher judgement is itself patterned, and that the patterns track language, gender and class. If that is so, it is an argument for putting the contradicting evidence in front of her in a form she can examine and reject. It is not an argument for moving the decision away from the only adult who knows the child, sees the home situation, and will be standing in front of the parent in April. Responsibility without authority is the fastest way to lose a good teacher’s engagement entirely.

01  What follows from it 4 demands
  1. 01.1

    Override without justification

    A teacher should be able to reject any conclusion a system reaches about a child without explaining herself, in no more steps than accepting it takes.

    Count the taps in each direction. If they differ, the design has an opinion about who should win.

  2. 01.2

    Disagreement is not a compliance failure

    A teacher who disagrees with a conclusion about a child is telling the supplier something about the conclusion, and a school should refuse an arrangement in which she is instead telling her head teacher something about herself.

    Ask whether override rates appear in any report a head teacher can see about staff. If they do, this is a surveillance loop wearing a support loop’s name.

  3. 01.3

    No silent re-assertion

    A school should be able to require that once a teacher has changed a decision, nothing reverses it later without that being visible and attributed.

  4. 01.4

    Every claim has an author

    A parent asking who decided this should receive a name, not a description of a process.

Index

An inference about a child must be answerable to that child’s parent

Two things could be meant by explanation and only one matters here. A parent does not need an account of how a system works. A parent needs the grounds: what the child actually did, when, and what was read from it. That distinction is what makes the requirement achievable without publishing a mechanism, and also what makes it demanding, because grounds can be checked and mechanisms cannot.

An explanation that cannot be checked is a reassurance. Telling a family their daughter is a visual learner, or that she is at risk, gives them nothing to inspect and nothing to contest. Telling them what their daughter actually did, and what was read from it, gives them something they can argue with. The second statement can be wrong in a way a parent can spot. That is the property worth having.

Language is not a presentation layer here. The parent may not read English, may not have finished school, and may be reading on a shared phone at the end of a working day. An explanation that exists only in the medium of instruction puts the child in the position of translating an assessment of herself to her own parents, which is unreliable and unfair to her. If an inference cannot survive being stated plainly in the home language, it was never clear enough to act on. And where the grounds are thin, the honest move is to make no claim at all. Saying nothing about a named child is a cost a school can weigh; saying something that cannot be shown to her parent is not.

02  What follows from it 4 demands
  1. 02.1

    Grounds, not mechanism

    A parent is entitled to see the work a claim about their child was drawn from, and to see it at home rather than across a desk in April.

  2. 02.2

    Written for the parent

    Explanations should exist in the languages families actually use, at a reading level that does not require the child to interpret an assessment of herself.

  3. 02.3

    No ungrounded claims

    Where the evidence will not support a statement about a particular child, a supplier should say nothing rather than something hedged.

    Silence has a cost of its own: a family told nothing learns nothing until the marks arrive. Which failure is worse, and for which families, we have not settled.

  4. 02.4

    Contestable by design

    A dispute should reach a person who can change the record, rather than a form that records that it was disputed.

Index

Fluency is not correctness, and a confident wrong answer costs more than no answer

A system that produces well-formed prose is optimised, by construction, for the appearance of knowing. It produces the register of the textbook, the shape of a derivation, the cadence of an explanation that has been checked. None of that is evidence about the content. This matters more in a classroom than almost anywhere else, because the recipient is a novice, and a novice is defined by being unable to tell a correct explanation from a fluent one. A student who could catch the error would not have needed the help.

The cost is asymmetric and usually described too weakly. A wrong answer does not merely fail to teach; it teaches. It installs a misconception with the authority of the school behind it, at the moment the student is most receptive, and that is harder to remove than the ignorance it replaced. Underneath sits a worse lesson: that confidence and polish are the marks of truth. Examinations punish that habit, and so does everything after them.

Refusal therefore has to be an ordinary, visible, unembarrassed state, and a system used with children should decline more often than one used with adults. Refusal reduces usage, and usage is what products are measured on, including ours. We take usage to be the wrong objective: a tool used less and trusted properly is better than one used constantly and believed uncritically.

03  What follows from it 4 demands
  1. 03.1

    Refusal is a normal state

    A school should be able to require that declining to answer is put to a child as an ordinary outcome, and should be suspicious of a supplier who treats it as a defect to be removed.

  2. 03.2

    See what the class was told

    A school should be able to see what its students were told, so that a wrong answer given to a whole class is corrected once rather than discovered thirty times.

  3. 03.3

    Working is a claim, not a proof

    Whatever working a child is shown should be taught to her as something to check, and never presented as the reason the answer is right.

    Worked steps can be fluent and wrong in exactly the way an answer can. Showing working is a help, not a safeguard.

  4. 03.4

    Assume the confident error

    Plan on the assumption that some students may have received a plausible wrong answer, and keep a routine for surfacing it.

Index

The effort that does the teaching must not be outsourced, and nobody knows yet which effort that is

The crude version of this position is that students should not have machines do their homework. It is too crude to use. The pupil who copies a solution has learnt nothing; the pupil who spends the hour on arithmetic slips during a lesson about forces has also learnt nothing about forces. Every task contains effort that builds the capability being taught and effort that is only the price of entry, and removing the second is one of the few unambiguously good things a machine can do in a classroom.

The difficulty is that the boundary is not a property of the task. It belongs to the pair: this task, this child, this week. Rewriting a messy solution neatly is the entire learning for a student who cannot yet organise an argument, and pure transcription for the girl beside her who organised it in her head twenty minutes ago. Look up a formula and you have either removed a distraction or removed the retrieval practice that was the point.

So we state this as an open problem rather than a policy. We do not know how to locate the load-bearing effort in a task reliably, for an individual child, across subjects, and we are not close. Anyone claiming to have solved it should be asked what result would have shown them wrong. What follows for a school is modest and firm: a line drawn once, for everyone, and left unstated is the wrong shape, and the people entitled to move it are the teacher and the child in front of her.

04  What follows from it 4 demands
  1. 04.1

    Ask where the line falls

    A school should be able to ask, of any task it sets, what a tool will do for a child and what it will leave to her, and to have the answer before the task is set.

  2. 04.2

    Decided for a child, not a class

    A teacher should be able to decide how much help a particular child gets, and not only how much the class gets, because the right amount differs within one room.

  3. 04.3

    Help can be withdrawn

    A teacher should be able to take the help away for a task, so she can find out what a student can do unaided when it matters.

  4. 04.4

    Know what was done unaided

    A teacher is entitled to know which part of a piece of work the child did herself.

    This is for teaching, not policing. Used to catch children rather than to read them, it will be defeated within a term.

Index

Evidence comes before scale, and most of the evidence was gathered somewhere else

We have not conducted a systematic review of the evidence on AI and learning, and this page is not one. What can be said without a review is structural: any result a supplier shows a school was obtained in some other classroom, because it was not obtained in this one. That is not an accusation, it is arithmetic, and it puts transferability on the buyer rather than on the seller. A result from a class of twenty may hold in a section of fifty-five sitting boards in a second language on shared devices, with a syllabus that must be finished by February. It may not. Nothing in the result itself settles which, and the school is the only party in the room with an interest in finding out.

Four questions cost a school nothing to ask. On whom was this measured, described precisely enough to compare with our children. Compared with what, meaning what the other group received rather than whether one existed. Who paid, and who chose what was reported. And the question that separates research from marketing: what result would have counted as failure. If no outcome would have been reported as a failure, nothing was tested, and what is on offer is a hypothesis with a price.

The standard applies to us, and by that standard we have published nothing. This site carries positions and questions. Where a reader expects a number there is none, and the omission is deliberate rather than coy: a figure detached from its population, its comparison and its failures is a decoration, and an accurate decoration is still not evidence. When work here reaches a point where population, comparison and negative findings can be published together, that is the form it will take.

05  What follows from it 4 demands
  1. 05.1

    Ask on whom

    Require a description of the students a claim was measured on, specific enough to compare with your own.

  2. 05.2

    Ask compared with what

    Establish what the comparison group received.

    Compared with nothing is not a comparison. A class given new attention of any kind may do better for a while for that reason alone, which is why what the other group received has to be stated.

  3. 05.3

    Ask what would have counted as failure

    A supplier who cannot describe a result that would have stopped the product has not run a study.

  4. 05.4

    Pilot in one room first

    Run anything new in a single section, with a teacher willing to say it did not work, before it touches a timetable.

Index

An approach that works only in English is not an approach for Indian schools

The ordinary situation in an Indian classroom is that instruction happens in one language, thinking happens in another, and the explanation that reaches the parent happens in a third. Children reason in the language they grew up in and write in the language of the examination, and there is loss in that passage. The loss is a translation gap and it is systematically mistaken for a knowledge gap.

That mistake is a mechanism by which capable children are sorted downwards. A system reading only what a student wrote in English records a vocabulary failure as a physics failure, and if that record then decides what she is given next, she is moved into easier material for a reason that has nothing to do with physics. Two states look identical on paper and need opposite lessons: the child who does not know the concept, and the child who knows it and cannot yet say it in the register the examination demands.

Translation does not settle this. Rendering a question into another language changes the question, because subject vocabulary often has no settled equivalent and because examination language is a dialect that has to be learnt eventually; removing it from a child’s path is deferral, not kindness. Mixed-language work, a sentence begun in one language and finished in another, is ordinary speech and not an error. A system that treats it as a mistake is telling a child that the way she thinks is a defect.

06  What follows from it 4 demands
  1. 06.1

    Separate language from subject

    A school should be able to require that a child’s difficulty with the language of the examination is never recorded as a failure to understand the subject, and should ask any supplier how it tells the two apart.

  2. 06.2

    Mixed-language work is valid

    A child who answers in the mixture of languages she speaks in should not be marked down for the mixture itself.

  3. 06.3

    Demonstrate in your languages

    Require a live demonstration in the languages of your classroom, on your own students’ work in the form your classroom actually produces it, before anything else is discussed.

    A demonstration in English proves nothing about anything else, including the same system one language over.

  4. 06.4

    Keep examination register in view

    Support in the home language is a route into the language of the examination, not a substitute for it.

Index

The weakest quartile is the test, not the average

An improvement in the class average is entirely compatible with harm. If a tool helps the confident students and confuses the students already behind, the mean rises, the spread widens, and the summary reports success. This is the expected failure mode rather than a rare one, because almost every tool asks something of its user, and the students best able to supply it are the ones who needed it least.

The lowest group is hardest to serve for structural rather than motivational reasons. They have the least prior knowledge with which to catch a wrong answer, so fluent errors land on them hardest. They read slowest, so interfaces built for comfortable reading exclude them first. They have the least reliable access to a device, so anything requiring continuity outside school fails for them specifically. And they complain least, so their difficulty appears as quiet non-use rather than as a defect report.

The position is that a tool is judged on whether the bottom of the room moved, and that a tool raising the mean while the bottom falls further behind should be called a failure. Against this it can be argued that total learning is what matters, and that a large gain concentrated among the students able to use it is still a large gain. A school is not a portfolio, though. It exists in large part for the children who would not get there otherwise, and those children remain in that room for years afterwards.

07  What follows from it 4 demands
  1. 07.1

    Report by prior attainment

    Any result about a tool is reported separately for the students who began at the bottom, or it is not reported.

  2. 07.2

    A widening spread is a result

    Treat an increase in the distance between the strongest and weakest students as a finding about the tool, not as background noise.

  3. 07.3

    Watch the quiet abandonment

    Track which students stop using something and stay stopped.

    The students who quietly give up are the ones a tool is usually justified by, so a fall in use among them is a finding about the tool rather than an absence of one.

  4. 07.4

    Design to the slowest reader

    A tool the weakest reader in the section cannot use without adult help is a tool for part of the class, and should be timetabled as such.

Index

Data minimisation is a design constraint, not a compliance exercise

The compliance question asks what a school can be brought to consent to. The design question is harder and more useful: what is the smallest amount of information about this child sufficient for the decision being made now, and how soon can the rest be released. The first produces a consent form. The second produces a different system.

Children’s records are not ordinary personal data. The person described is a minor, the consent was given by someone else, and the record outlives both the child’s time at the school and the school’s relationship with the supplier. Something written about a thirteen-year-old having a bad term should not still be a fact about her at twenty-five. She had no practical means of contesting it when it was written and will have none later, because by then it will be a number in a summary whose origin nobody remembers.

The cost is real and we accept it. A broader record about a child would probably support better inferences, and holding less means accepting worse ones. We take that trade for children, and the argument against it is a serious one: that a more accurate system helps a child more than a leaner one protects her. It should be had in those terms rather than settled by a checkbox.

08  What follows from it 4 demands
  1. 08.1

    A purpose for every item

    Ask what decision each item held about a child is required for. Anything kept because it might be useful later is building the wrong system.

  2. 08.2

    Deletion demonstrated, not promised

    Ask to watch a record be deleted, then ask what is retained elsewhere.

    Removed from the interface is not the same as deleted, and the difference should be given to the school in writing.

  3. 08.3

    No protected characteristics inferred

    Caste, religion, income, family circumstance and health are not inferred from behaviour, whether or not the inference would be accurate.

  4. 08.4

    Know what leaves the school

    Establish which data leaves the premises, who else sees it, and what happens to all of it on the day the contract ends.

Index

Assessment should change Monday’s lesson, not only record July’s result

The case against the mark is set out on the position page and is not restated here. What this position needs from it is narrow. A mark is a reasonable instrument for ranking and a poor one for teaching: it identifies who is behind, and it does not tell a teacher what to do with the forty minutes she has tomorrow, in front of fifty-five children, with a syllabus that must be finished on time.

The useful output of an assessment is therefore an action rather than a score. Reporting that a chapter went badly is not a lesson plan, because a chapter is not something anyone can teach in a period. Reporting a specific misunderstanding that a teacher could take up in one period is. Timing belongs to the instrument and not to operations: an assessment arriving after a unit has been taught cannot change how it was taught, only inform a repetition the calendar does not allow.

Two arguments run alongside this position rather than against it. The first is that a class taught only what it got wrong receives a fragmented subject, assembled from its own errors and with no shape. Diagnosis is an input to planning and not a replacement for it, and a teacher who chases a list of misconceptions will produce students who can patch errors and cannot see the whole.

The second is the argument this site makes against the mark, turned on whatever is proposed in its place. A measurement that compresses deforms the curriculum, and it does not follow that a measurement which compresses less is safe. Any account of a learner more detailed than a mark, used inside a system that allocates scarce seats, offers more surfaces to teach to rather than fewer, and more places where a record can be curated before anyone reads it. It may deform teaching worse than the mark does, precisely because it is more specific about what to deform. We do not know how to prevent that. Three things would tell us it had started: a school reporting such an account upwards instead of teaching from it, a teacher appraised on what it says about her section, and a parent who asks to have it improved rather than asking whether it is true.

09  What follows from it 4 demands
  1. 09.1

    Output is a teaching action

    An assessment result should name something a teacher could actually do in one period.

  2. 09.2

    Reasons over counts

    A count of errors does not tell a teacher what to do; the shape of the error might. A school should ask a supplier what its results say about how a child went wrong, and not only how often.

  3. 09.3

    Arrive before the unit ends

    Assessment intended to change teaching is scheduled while the unit is still being taught.

    If the only feedback loop runs from examination to examination, nothing inside it can change instruction, however good the analysis.

  4. 09.4

    Diagnosis does not replace planning

    Keep the sequence of the subject intact; use diagnosis to decide where to slow down, not what the subject is.

Index

Some decisions should be kept away from these systems entirely

This is not a call for general caution. It is a list, and a school can adopt it without buying anything or waiting for anyone’s research.

Irreversible sorting comes first: streaming, gatekeeping, selection, exclusion, and any prediction of how far a child will go, which we expect to arrange its own confirmation, because the adults who see it teach accordingly and the child eventually hears it. Second, inference about a child’s inner life: emotional state, motivation, character, family circumstance, mental health. Defensible across a population, such an inference becomes, when attached to one named child, an unfalsifiable claim a school will nonetheless act on. Third, discipline. Fourth, unsupervised communication with a child, which schools already know how to think about and should not stop thinking about because the counterpart is a system. Fifth, public comparison of children against one another, a bad practice before any of this and not improved by automation.

The first item carries a real cost. Identifying children at risk of dropping out is how a school helps them, and refusing the prediction has a cost of its own, paid by children nobody noticed in time. The same flag that marks a child for support also marks her for lowered expectations, and in a school under pressure, with too many children and too few hours, we would expect the second reading to be taken more often than the first. The objection is therefore not to the prediction but to the institution receiving it. Where a school can show that a flag reliably produces extra teaching rather than reduced ambition, the position should be revisited for that school.

10  What follows from it 5 demands
  1. 10.1

    No automated gatekeeping

    Placement, selection, streaming and exclusion are decided by people who can be asked to justify them, with the system contributing evidence and not conclusions.

  2. 10.2

    No inference about inner life

    Emotional state, motivation, character and home circumstances are not inferred from a child’s interactions with a system.

  3. 10.3

    No unsupervised contact

    An adult can see everything said to a child, and this is not traded away for convenience or scale at any point.

  4. 10.4

    No public ranking of children

    Comparative displays of children against each other are not produced, whatever the intended motivational effect.

  5. 10.5

    Escalation stays human

    Any signal that a child may be at risk goes to a named person, not into an automatic action.

    Ask the harder question first: when a flag is raised in your school, does the child receive more teaching or lower expectations? The honest answer decides whether the flag should exist.

Standing

What this page is, and what it is not

  1. These are positions, not findings. We are prepared to defend them and to lose them. Where a position rests on an empirical question we cannot answer, the page says so rather than borrowing confidence from elsewhere.

  2. No results, benchmarks or effect sizes appear here, and nothing about how any system is built. A number without its population, its comparison and its failures is a decoration; a mechanism described in public is a claim nobody can check.

  3. We do not meet every demand on this page in everything we build today. That is a gap in what we have built and not in the standard, and a school is entitled to ask us which items those are before it asks anyone else.

  4. Research asks what is possible, the Council is meant to decide what is responsible, and the products show what can be built. This page belongs to the first and constrains the third. Its accountability to the second is a commitment rather than a protection a school can rely on today: the Council’s research instrument is a founding draft, published on 29 August 2026 and open for comment, the Council has no members, no ethics review panel has been appointed, and no study has begun.

  5. This page was published on 29 August 2026 and has not been revised since. When the argument against a position here wins, the position will be withdrawn on this page, with the date and what replaced it. Disagreement, particularly from teachers running sections of fifty and above, is the intended response.

Positions, not a standard

Ten positions this company holds. No one has adopted them.

Nothing certifies these, no body has adopted them, and they carry no authority beyond the argument made for each one. A school is free to take four of them into a meeting and leave the other six here.

Where one of them is wrong, the useful thing to do with this page is to say so — and the questions we expect, including the uncomfortable ones, are answered separately.

Disagreement is the intended response. Write to us.