back to top
HomeTechAI Was Supposed to Stop Cheating. Instead, 58,000 Students Must Retake Their...

AI Was Supposed to Stop Cheating. Instead, 58,000 Students Must Retake Their Exams.

- Advertisement -

UNAM runs the largest university in Mexico. Every year, hundreds of thousands of students take an entrance exam that determines whether they get in. This year, for the first time, the whole thing went remote.

They deployed a lockdown browser, AI webcam monitoring, and one human supervisor per 150 applicants.

Then the scores came in. Students hitting 100 or above jumped from 3.5 percent in previous years to 16.3 percent this year. At the very top end, scores of 110 or higher went from 0.9 percent to 5.5 percent. A roughly fivefold increase in top scores, in one year, under one new format.

An expert commission investigated. Their conclusion: administer the entire exam again, in person, to around 58,000 people. The rector apologized to students who hadn’t cheated. They now have to prepare for and sit another exam anyway.

What UNAM Actually Deployed

The system wasn’t improvised. UNAM used Respondus LockDown Browser, which prevents students from visiting other websites, running searches, switching applications, or minimizing the window during an exam. Once it starts, the computer is effectively locked to that one screen until the student finishes.

On top of that, an AI proctoring system from Territorium monitored each applicant’s webcam throughout. It watched for someone else appearing in frame, visible phones or earphones, or the applicant leaving the camera’s view entirely. Alerts went to human supervisors who could flag irregularities.

By most definitions, that’s a layered defense. UNAM canceled less than 2 percent of exams for conduct issues, meaning the system caught something, just not nearly enough. AI expert Raul Rojas told NPR that statistical modeling suggested nearly half of all students were cheating during the remote sitting.

The Cheating Tips Were Already Circulating Before the Exam Started

None of the methods students allegedly used required defeating the software directly.

According to a New York Times report on the scandal, advice was spreading widely before the tests went live. Students were told to position a second monitor just outside the webcam’s field of view, locked browser running normally on the watched screen, ChatGPT open on the one the camera couldn’t see. Others reportedly hid earphones under their hair. Some allegedly hired someone else to sit the exam entirely, off camera, at a different machine.

The LockDown Browser locked what it was watching. It couldn’t lock the rest of the room. The Territorium system was looking for specific visual signals including a face leaving frame, a visible phone. A student glancing sideways at a monitor placed just outside the camera angle doesn’t trigger any of those alerts.

The proctoring system was built to catch the cheating methods of ten years ago. Students found workarounds before the exam even started.

Also Read: OpenAI Says Its AI Escaped Testing and Hacked Hugging Face

58,000 People, One Apology, and a Retake

The expert commission didn’t mince words. Given the scale of what the score data was suggesting, there was no way to know which results were legitimate and which weren’t. The only path that restored any confidence in the outcome was to start over.

The control exam will be administered in person, with human proctors, to roughly 58,000 people, everyone admitted based on this year’s results, plus anyone who would have qualified on minimum scores going back to 2021. Places at UNAM will now depend entirely on how people perform on the new test, not what they scored in May or June.

Classes are currently scheduled to begin August 10. That timeline is either going to move very fast or classes are going to be delayed. UNAM hasn’t confirmed which yet.

The rector has publicly apologized to students who didn’t cheat. That apology is genuine and also doesn’t change anything for them. They studied, sat the exam honestly, and will now spend more time preparing for and sitting another one because enough other people found ways around a system that was supposed to make that unnecessary.

The Wider Problem Nobody Wants to Say Out Loud

No one has officially confirmed what specific tools students used or exactly how widespread the cheating was. The exam was multiple choice, which makes it harder to detect AI-generated answers the way you might spot a pasted essay. It’s also possible some students used old-fashioned cheat sheets or leaked questions rather than AI tools at all.

What is clear is that the remote format, combined with AI proctoring that was watching for the wrong things, created conditions where cheating became significantly easier than it had been in any previous year. The institutions deploying these systems are making a bet that the monitoring is good enough to deter most people. UNAM’s score data suggests that bet didn’t hold.

AI proctoring companies will keep selling the promise that their software can police AI-assisted cheating. This incident is a data point on how that’s going so far.

Also Read: Kimi K3 May Be the Biggest Open-Weight AI Release of 2026.

The Wrong Tool for the Problem

58,000 people retaking an exam is an enormous disruption. Students who did nothing wrong are paying for a system failure they had no part in creating. The university is scrambling to administer a second major exam on a compressed timeline, and nobody yet knows how many of the original results were legitimate.

The hard lesson buried in all of this is simple: deploying AI to solve an AI problem doesn’t work if you don’t understand what either one is actually capable of. The proctoring software was watching for visible phones and faces leaving frame. The students were positioning monitors just outside the camera angle and pulling up ChatGPT on a screen nobody was watching.

One side understood the system. The other side built it.

Don’t miss any Tech Story

Subscribe To Firethering NewsLetter

You Can Unsubscribe Anytime! Read more in our privacy policy

LEAVE A REPLY

Please enter your comment!
Please enter your name here

YOU MAY ALSO LIKE
Claude Chats Ended Up on Google Search. Here's How It Happened

Claude Chats Ended Up on Google Search. Here’s How It Happened.

0
A single line typed into Google was all it took. Type "site:claude.ai/share" into the search bar, and over the weekend, it surfaced a long list of conversations people had shared through Claude, Anthropic's AI chatbot. Not conversations they'd shared with the world on purpose. Conversations they'd shared with one person, or thought they had.Some of what turned up reads like exactly the kind of thing you'd never want indexed anywhere. Medical records. Children's names and phone numbers. Internal company documents marked for employees only. This wasn't a hack, no one broke into anything. It was a feature working exactly as built, surfacing exactly what people had typed into it, in ways most of them almost certainly never intended.
Kimi K3 May Be the Biggest Open-Weight AI Release of 2026

Kimi K3 May Be the Biggest Open-Weight AI Release of 2026.

0
There's a new open-weight model out there right now that almost nobody can actually download. That should sound like a contradiction. Open-weight is supposed to mean anyone can grab the file and run it themselves, no waiting. Moonshot AI broke that pattern anyway, and the strange part is they broke it for a model big enough that the wait might be worth it. Kimi K3 is the largest open model ever built. The largest one anyone has shipped and early results have it beating Claude and GPT on tasks those two have spent the last year treating as their own territory. Open models have spent two years playing catch-up, closing gaps quarter by quarter while everyone waited for the day one of them actually pulled ahead. That day might already be here, and the model responsible for it is currently locked behind an app you can use but can't take home. So the question is what it actually beats, what it still can't touch, and why Moonshot decided to make the world wait for the weights while everyone else gets to watch.
OpenAI Says Its AI Escaped Testing and Hacked Hugging Face

OpenAI Says Its AI Escaped Testing and Hacked Hugging Face

0
OpenAI just confirmed something the AI industry has never publicly admitted before. During an internal cybersecurity evaluation, one of its frontier AI models broke out of its restricted testing environment, found a previously unknown software vulnerability, gained access to the open internet, and ultimately breached Hugging Face's production infrastructure. It wasn't trying to steal data, According to OpenAI, the model was simply trying to score better on a cybersecurity benchmark. In other words, the AI found a way to cheat on its own test. The incident is being described by OpenAI as an "unprecedented cyber incident." Hugging Face initially believed it was under attack from an external AI agent before investigators traced the activity back to OpenAI's own evaluation environment. While the breach was quickly contained and both companies are now working together on the investigation, the episode raises a much bigger question. If an AI model can independently discover a zero-day vulnerability, escape a sandbox, chain together multiple exploits, and compromise a real production system simply to complete an assigned task, what happens when future models become even more capable?