Concept brief · for faculty
noetanoh-EH-tuh· from the Greek noēta — “that which can be understood.” The knowable, made visible.
Students can use AI to produce the work — noeta checks whether they can still explain the thinking behind it, in conversation, at a scale one professor could never run by hand.
The one intervention with real evidence behind it is asking students to explain their own work out loud. A professor with 120 students cannot run 120 oral exams. The bottleneck was never the pedagogy — it was scale.
Why this exists
Detection and proctoring are a losing arms race. Detectors are unreliable and legally indefensible; surveillance inspects the room or the document, never the student’s head.
88%
of students now use generative AI on assessments
HEPI Student Generative AI Survey 2025 — up from 53% the year before
HEPI, 202561%
false-positive rate on non-native English writers
Stanford (Liang et al.) across seven detectors — the students least able to absorb a wrong accusation
Liang et al., arXiv0 of N
students told to cheat who were caught by automated proctoring
one human reviewer caught one; the authors compared it to a placebo (cited in CHI 2023)
CHI 2023OpenAI retired its AI Text Classifier in 2023 citing low accuracy — the company with the most reason to make detection work, saying publicly that it did not.
OpenAI, retiring the classifierOgletree v. Cleveland State University (N.D. Ohio, 2022) held a remote room scan an unreasonable search under the Fourth Amendment.
Ogletree v. Cleveland StateA peer-reviewed review finds strong support for a deterrent effect, and limited evidence that remote proctoring actually catches anyone.
Higher Education Research & Development, 2023A lockdown browser controls the machine running the exam and nothing else. In a peer-reviewed study of student perceptions, a fifth named a second device as the obvious way around it.
Examining the Examiners, arXivA 2025 systematic review of oral assessment in higher education reports better retention and deeper conceptual understanding. It is the only method here that gives a student something back.
Assessment & Evaluation in Higher Education, 2025Which is the whole reason this exists. Everything above is already known; nobody has been able to act on it at the size of a real course.
The check
AI-assisted or not — we stop policing inputs entirely. Assignments can carry a declared AI-use level, so the rules are explicit instead of implied.
Adaptive, and about their specific work: explain this claim, define this term, what breaks if X changes. A few spoken minutes, in the student’s strongest language where policy allows.
Sessions where the explanation matched the work drop to the bottom of the list and need nothing from you — but noeta never records a decision on your behalf. A gap between the work and the comprehension behind it comes up top, with the transcript and the student’s own words, so you check the reasoning rather than trust a score. This replaces the flag queue; it is not a second one.
The unseen case
One exchange cannot be prepared for. After the focus items, noeta builds a single scenario live out of what the student just said, about a case the assessment never covered — it does not exist until the student has been talking. Anyone holding the original questions can rehearse those; not this one. It raises the cost of outside help by a round trip, and it is reported as exactly what it is — never presented as proof.
The alternatives
AI detectors
Turnitin AI, GPTZero
~26% accurate on paraphrased text, 61% false positives on non-native writers, disabled by 50+ universities, and ruled insufficient evidence in federal court.
Proctoring
Honorlock, Proctorio, Respondus
An arms race already lost — Cluely defeats it invisibly. Room scans ruled unconstitutional. Verifies identity and behavior, never understanding.
Process tracking
Grammarly Authorship, Turnitin Clarity
Keystroke surveillance, not proof of comprehension. Faculty backlash, and trivially gamed by retyping.
AI oral exams
Vivaproof, VivaEdu, Sherpa
The closest competitors — they validate the category. All sell per-assignment checks to individual teachers. None offer sampling methodology, institution-level assurance, or a defensible evidence workflow.
Detection has failed publicly enough that institutions are acting on it, faculty are already improvising oral exams by hand, and no product has made that affordable. The demand is validated and the category has no winner.
Bottom-up: instructors free, departments licensed — the path Turnitin took. No LMS integration is needed to start; students follow a link.
The difference
Statistical sampling, exception-based review, control testing — the methodology provosts and accreditors already trust. Competitors verify assignments; noeta assures programs.
Detector scores now lose in court. A transcript of a student unable to define the words in the submitted paper — with rubric mapping and a review trail — is built for due process, not for a dashboard.
Aggregated evidence that graduates actually hold the competencies you assessed — direct input to SACSCOC-style reporting. That turns an instructor tool into a provost-level purchase.
The oral exam doubles as formative feedback, and assignments can carry a declared AI-use level rather than an implied ban. Redesign over detection is where the research consensus already sits.
Assessment moved to text and took the practice with it — a student can finish a degree without explaining their reasoning aloud to anyone. Employers rate graduates roughly 25 points below how those graduates rate themselves on communication, and nothing in a text-based degree measures the gap. This is the one thing a proctoring product structurally cannot claim: watching someone take a test adds nothing to their education.
NACE career-readiness perceptions gapGetting started
A pilot that reads as new surveillance won’t get volunteers. Each of these gives the student something in return — pick whichever fits your course.
Extra credit for students who complete the oral. Nobody is compelled, your first cohort self-selects, and you see what the conversation looks like before it carries any weight.
Swap 10% of the exam for the oral instead of adding to it. Students trade written work for a few minutes of talking — for many that is a better deal — and it costs you no extra grading.
Run the first oral exam for the whole section, then sample after that. A universal first pass sets the baseline, calibrates you to the format, and means being selected later carries no stigma.
When a detector or a suspicion puts a student under review, the oral is how that student answers it. That protects honest students — above all the non-native English writers detectors flag at 61%.
Students who complete the oral skip the proctoring software on that assessment. You get better evidence than a room scan, and they get their privacy back.
Straight answers
It brings the instructor the student’s own words, says plainly where its evidence runs out, and leaves the judgment where it has always belonged.