Benchmark · September 2026

How accurate is Silent Student?

Seven homework assignments, built as replicas of the platforms students actually get, worked start to finish and scored by reading the answers back.

In our latest benchmark, Silent Student answered 156 of 164 graded answers correctly (95%) across seven homework assignments built as replicas of WebAssign, Cengage MindTap, McGraw Hill Connect, Pearson MyLab and Macmillan Achieve, and submitted six of the seven. It ran on the same engine the Base plan uses.

95%graded answers correct (156 of 164)
6 of 7assignments submitted
9 of 12deliberate faults recovered from

Results by assignment

AssignmentReplica modeled onWhat it testsCorrectSubmitted
STEM homework workbookWebAssignPer-part answer checks, retries, figures and a required submission note8 of 8Yes
Mixed assessmentCengage MindTapTwo-pane activity player, custom pickers, one attempt per question9 of 9Yes
Accounting workbookMcGraw Hill ConnectDense multi-cell grids, requirement tabs and a question map, 109 graded parts102 of 109Yes
Calculus homeworkCengage WebAssignGraphs as answer choices, 3D figures, a 10-part accordion, math notation16 of 16Yes
Finance homework (first third)Pearson MyLabCustom pickers, per-part checks, three tries per part9 of 9Yes
Statistics assignment (first third)Macmillan AchieveSorting bins, numeric and short answers, image choice7 of 8No — stopped on one question
College algebra (first third)Pearson MyLab (College Algebra)Math editor inside nested frames, immediate feedback5 of 5Yes
Total156 of 164 (95.1%)6 of 7

How the benchmark works

  • Replicas, not live accounts. Each assignment is a local replica built to behave like the real platform: the same question types, answer controls, checks and submission flow. No real student account or live vendor site was used.
  • Deliberate faults. Four of the assignments sometimes reject a correct answer or make a control disappear, to test whether the bot recovers instead of giving up. It recovered from 9 of the 12.
  • Independent scoring. After submission, each replica is read back to see which answers it recorded as correct. The bot’s own report is not used.
  • Same engine as the app. The run used the model and settings the Base plan uses for platforms without a dedicated solver. In the app, real WebAssign, MindTap and SmartBook assignments go to dedicated solvers instead.
  • When. One complete run of all seven assignments on 22 September 2026.

What this doesn’t show

It is one run of one engine on seven assignments. Your course, question types and platform version may differ, live platforms change, and some question types are harder than others. It says nothing about proctored exams, which Silent Student does not take.

Were these real WebAssign, Pearson or McGraw Hill accounts?

No. Each assignment is a replica we built for testing, so no real student accounts or live vendor sites were used. The Pearson Finance and Achieve replicas are authorized one-to-one captures of the first questions of real assignments; the others are clean-room reconstructions of how those platforms behave.

How was each answer scored?

After each assignment was submitted, the replica itself was read back to see which answers it recorded as correct. The score does not rely on the bot's own report.

Does this benchmark predict my results?

It measures one engine on seven assignments. Your course, question types and platform version may differ, and some question types are harder than others.

Last updated 29 September 2026. See also security and privacy and supported platforms.