Fall syllabi are starting to read like agent-evaluation specs. Universities publishing AI-use rubrics for the new term are spelling out when autonomous tools may retrieve sources, draft feedback, or stay barred from assessments—and who signs off before students see grades.

Rubrics beyond plagiarism

Oregon State’s partnership with Metrum AI turns presentation videos into evidence-tagged draft reviews aligned to faculty rubrics, with instructors retaining final authority. Finance professor Jonathan Kalodimos, who prototyped the system, describes “rubric engineering” as structured extraction of performance signals rather than a single AI score.

Appalachian State’s veterinary technology pilot uses a grading assistant to flag rubric milestones in video transcripts, but faculty must validate skills. The pattern is consistent: agents accelerate triage; humans own outcomes.

Teaching students to evaluate agents

Stanford’s CS329Z course on engineering AI agents dedicates weeks to evaluation harnesses—defining requests, environments, stopping criteria, and scorers. Homework asks students to build benchmarks with code graders and model-based judges, mirroring industry practices as homework bots proliferate undergrad courses.

HKUST’s RubriX project goes further, using multi-agent workflows so students and instructors co-create rubrics for writing assignments across engineering and humanities sections. The goal is transparency about how criteria are interpreted, not hidden automation.

Industry debate lands on campus

Meta CEO Mark Zuckerberg argued Tuesday that AI labs should rely on independent evaluators and market incentives rather than an industry-wide capability slowdown, echoing Nvidia’s Jensen Huang. He noted Meta delayed its Muse agent for months over safety reviews while still shipping consumer features.

Campus administrators hear that tension directly: trustees ask whether procurement should demand third-party red-team reports, while faculty want academic freedom to experiment. Published rubrics are the compromise—public criteria agents must meet before assisting with instruction.

Implementation headaches

Accessibility offices insist approved assistive tech cannot be blocked by blanket agent bans. International students question whether translation agents require separate disclosure. Libraries host clinics on citing model assistance, similar to citation managers a decade ago.

Accreditation visits this fall will likely ask how programs prove graduates can perform without agents on high-stakes exams, even as low-stakes homework allows guided use. Clear rubrics, backers say, beat silent assumptions that led to honor-code disputes last year.

Whether Zuckerberg’s evaluator-centric vision satisfies critics of frontier labs, universities are already building the scoring frameworks agents will face in classrooms long before regulators settle on federal rules.

Student government associations at several public flagships asked provosts to publish rubric templates centrally so transfer students see consistent labels across departments, reducing confusion when every syllabus uses different agent vocabulary.