Search
Results for research
-
When AI Beats Economists – And Why That’s Good News For Assessment
Not so long ago, the idea that an AI system could out‑analyse a room full of economists would have sounded like science fiction. Yet that’s exactly what a recentFederal Reserve working paper set out to test.
-
Can Foundation Models Really Replace Human Judgement? What GPT-5.2's Launch Today Tells Us, and Why RM Compare Matters More Than Ever
Today, December 11, 2025, OpenAI launched GPT-5.2, its "most advanced model yet for professional work," with claimed improvements in coding, reasoning, long-context analysis, and fewer hallucinations than any previous version. The benchmarks are impressive: GPT-5.2 Thinking scored 100% on a challenging mathematics test, outperformed human professionals on 70% of well-defined knowledge work tasks across 44 occupations, and reduced factual errors by 30% compared to its predecessor.
-
Who is Assessing the AI that is Assessing Students?
As AI steps into the heart of education, we celebrate the speed and efficiency of machine-marked assessments. But a deeper question shadows every advance: If an AI can now judge student work, who—if anyone—is judging the AI?
-
'Cognitive Offloading', 'Lazy Brain Syndrome' and 'Lazy Thinking' - unintended consequences in the age of AI
Generative AI and LLMs are everywhere in 2025 classrooms, making teaching and learning faster than ever. But as research in the Science of Learning shows, letting technology do too much of the thinking for us—“cognitive offloading”—weakens the very skills education aims to build.
-
What the Recent Apple Study Taught Us About AI, Reasoning, and Assessment
In June 2025, Apple published a landmark study that has sent shockwaves through the AI and assessment communities. The research, titled The Illusion of Thinking, rigorously tested the reasoning abilities of the most advanced AI models—so-called large reasoning models (LRMs) from OpenAI, Google, Anthropic, and others—using a series of classic logic puzzles designed to scale in complexity. The findings have profound implications for how we understand AI’s capabilities, especially in the context of educational assessment.
-
RM Compare and AI
We get asked a lot of questions about RM Compare and AI - here are some of the common ones together with our responses
-
Fostering Critical Thinkers: RM Compare in the Humanities Classroom
In the Humanities, Welsh teachers face a unique challenge: how to effectively assess critical thinking, analytical skills, and the ability to construct well-reasoned arguments.
-
From Cheating Prevention to Assessment Validity: A Paradigm Shift in Digital Assessment
In the rapidly evolving world of educational technology, a groundbreaking study published in Assessment & Evaluation in Higher Education is challenging our approach to digital assessments.
-
Higher and Further Education
High profile Universities around the world use RM Compare to deliver student impact
-
RM Compare Newsletter February 2024