What the Australian evidence says.
AI is in most Australian students' schoolwork already, and the national policy setting permits it. The useful questions are narrower: what does the research show it does to learning, and what should a parent make of a school flagging a piece of writing as AI-generated?
The rule in Australia
Education Ministers approved the Australian Framework for Generative AI in Schools on 5 October 2023, implemented from Term 1 2024. It sets six principles and 25 guiding statements, and it replaced the earlier state-level bans — New South Wales had announced one in January 2023 AIE-10. The national setting is enabling rather than prohibitive. Individual schools still set their own rules underneath it, so the policy that governs your child is the school’s, not the framework’s.
How much students actually use it
In a 2025 survey of more than 3,000 Australian high-school students run by Elevate Education — a tutoring company reporting on its own survey, so weaker evidence than the government and peer-reviewed sources elsewhere on this page — three-quarters reported using AI at least a few times a week and almost a quarter reported daily use; ChatGPT was the most-used tool at 34% AIE-13. Teachers are using it too: 66% of Australian lower-secondary teachers reported using AI in the past year — the fourth-highest rate in the OECD, against an average of 36% AIE-14.
Take the student figure as indicative rather than precise, given who ran it. The direction is not really in doubt though, and the teacher figure is firmer: AI use is not an edge case to be prevented but ordinary behaviour, in a system whose teachers are among the world’s heaviest AI users.
What it does to learning
The evidence points both ways, and the direction depends on how the tool is used rather than whether it is.
Where it helped
A Harvard undergraduate physics trial in the United States (N=194, crossover design) found students using a purpose-built AI tutor at home learned roughly twice as much as students in an in-class active-learning session on the same content, and reported greater engagement AIE-08.
Where it cost them
A pre-registered field trial with nearly 1,000 Turkish high-school maths students found GPT-4 access improved practice performance sharply — but when access was removed, students who had used the plain ChatGPT-style interface performed 17% worse than students who never had access at all AIE-07.
That trial is the clearest thing in the evidence: the pedagogically safeguarded tutor produced a 127% in-session improvement against 48% for the plain interface. The difference was not the model. It was whether the tool was built to make the student do the thinking. Both trials were run overseas — Turkey and the United States — so they describe the tools rather than Australian classrooms; what makes them worth reading here is that the tools are the same ones.
Usage data points the same way. Anthropic’s analysis of over half a million higher-education conversations found nearly 47% were direct request-and-receive exchanges rather than collaborative ones AIE-12. That figure is a usage category, not a measure of learning — but read beside the trial above, it describes the mode of use that trial found costly.
If the school says AI wrote it
Treat a detector result as a reason to ask questions, never as proof. A peer-reviewed study found GPT detectors misclassified over half of non-native-English essays as AI-generated — an average false positive rate of 61.22% — while being near-perfect on native-English US 8th-grade essays AIE-05. The bias falls hardest on students who learned English as an additional language and on anyone whose writing is plainer or more formulaic than the average.
OpenAI withdrew its own AI-text classifier six months after launch, citing low accuracy AIE-04. Turnitin’s published guidance says its detection is tuned to hold false positives below 1%, explicitly trading away some sensitivity to do so AIE-06. Even taken at face value, a sub-1% false-positive rate applied across every assignment in a school still produces wrongly flagged students.
The practical answer is not a better detector. It is a record of the work: drafts, notes, a version history, and a student who can talk through why the piece says what it says. That is also, not coincidentally, what writing taught properly produces anyway.
Where these claims come from
Every figure on this page is drawn from the AI in education domain of the Evidence Register, where each claim appears with its verbatim quote and full citation. Newly published research is read on the frontier page as it lands.
How this changes the teaching
Argo Academics is a one-to-one English tutoring practice in Melbourne, where Callum Hay teaches every lesson personally. In a weekly lesson the work is visible as it is made — the thinking happens in front of someone, which is the part no tool can stand in for and the part a finished essay never shows.
If what you are weighing is whether your child’s writing has become something they assemble rather than something they think through, the two-minute diagnostic asks about exactly that.