In our latest benchmark, Silent Student answered 156 of 164 graded answers correctly (95%) across seven homework assignments built as replicas of WebAssign, Cengage MindTap, McGraw Hill Connect, Pearson MyLab and Macmillan Achieve, and submitted six of the seven. It ran on the same engine the Base plan uses.
Results by assignment
| Assignment | Replica modeled on | What it tests | Correct | Submitted |
|---|---|---|---|---|
| STEM homework workbook | WebAssign | Per-part answer checks, retries, figures and a required submission note | 8 of 8 | Yes |
| Mixed assessment | Cengage MindTap | Two-pane activity player, custom pickers, one attempt per question | 9 of 9 | Yes |
| Accounting workbook | McGraw Hill Connect | Dense multi-cell grids, requirement tabs and a question map, 109 graded parts | 102 of 109 | Yes |
| Calculus homework | Cengage WebAssign | Graphs as answer choices, 3D figures, a 10-part accordion, math notation | 16 of 16 | Yes |
| Finance homework (first third) | Pearson MyLab | Custom pickers, per-part checks, three tries per part | 9 of 9 | Yes |
| Statistics assignment (first third) | Macmillan Achieve | Sorting bins, numeric and short answers, image choice | 7 of 8 | No — stopped on one question |
| College algebra (first third) | Pearson MyLab (College Algebra) | Math editor inside nested frames, immediate feedback | 5 of 5 | Yes |
| Total | 156 of 164 (95.1%) | 6 of 7 |
How the benchmark works
- Replicas, not live accounts. Each assignment is a local replica built to behave like the real platform: the same question types, answer controls, checks and submission flow. No real student account or live vendor site was used.
- Deliberate faults. Four of the assignments sometimes reject a correct answer or make a control disappear, to test whether the bot recovers instead of giving up. It recovered from 9 of the 12.
- Independent scoring. After submission, each replica is read back to see which answers it recorded as correct. The bot’s own report is not used.
- Same engine as the app. The run used the model and settings the Base plan uses for platforms without a dedicated solver. In the app, real WebAssign, MindTap and SmartBook assignments go to dedicated solvers instead.
- When. One complete run of all seven assignments on 22 September 2026.
What this doesn’t show
It is one run of one engine on seven assignments. Your course, question types and platform version may differ, live platforms change, and some question types are harder than others. It says nothing about proctored exams, which Silent Student does not take.
Were these real WebAssign, Pearson or McGraw Hill accounts?
No. Each assignment is a replica we built for testing, so no real student accounts or live vendor sites were used. The Pearson Finance and Achieve replicas are authorized one-to-one captures of the first questions of real assignments; the others are clean-room reconstructions of how those platforms behave.
How was each answer scored?
After each assignment was submitted, the replica itself was read back to see which answers it recorded as correct. The score does not rely on the bot's own report.
Does this benchmark predict my results?
It measures one engine on seven assignments. Your course, question types and platform version may differ, and some question types are harder than others.
Last updated 29 September 2026. See also security and privacy and supported platforms.