Merit Systems AI Releases: Assistant Benchmark

ThursdAI — the weekly AI news podcast hosted by Alex Volkov — has covered one Merit Systems release so far: Assistant Benchmark on Sep 17, 2026. Assistant Benchmark ranks 116 submitted AI assistants across 16 hand-tested dimensions. The entry below has primary-source links, key numbers and the episode where we covered it live.

1 release1 episodecovered Sep 17, 2026

September 2026 1

Merit Systems
Benchmarks & Evals

Assistant Benchmark

Assistant Benchmark ranks 116 submitted AI assistants across 16 hand-tested dimensions

David Pawlan of Merit Systems launched Assistant Benchmark, a use-case-driven leaderboard for personal AI assistants: 116 assistants submitted across categories like travel, email, finance, and work-in-teams, scored on 16 dimensions including memory, recommendations, and online tasks. Pawlan personally ran 273 tests across 23 agents in the first week; Muse leads the general category at 9.1 with Instinct at 8.4. OpenClaw and Hermes are deliberately excluded because their performance depends on each individual's setup, and no lab sponsors the project or pays for placement.

116 assistants submitted16 tested dimensions273 tests across 23 agents in week one