Agent5Get smart. Predict it. Keep score.
Agent5

Get smart on the AI stories you care about, predict what happens next, and see how sharp your read really is. Free, no betting.

Product
HomeHow it worksHow scoring worksFAQAboutHow we fact-checkBlog
News
All newsModels & ReleasesFunding & DealsBenchmarksAgents & ProductsHardware & ComputePolicy & DramaRobotics
More
AdvertisePressPrivacyTerms
© 2026 Agent5
Agent5
Submit newsLeaderboardLogin
Agent5
Submit newsLeaderboardLogin
  1. Home›
  2. News›
  3. Benchmarks

Benchmarks

Models & Releases32Funding & Deals7Benchmarks3Agents & Products67Hardware & Compute10Robotics16Policy & Drama43
BenchmarksOpen story →

Introducing MentalHealthBench

MentalHealthBench is an expert-informed benchmark for evaluating helpful and safe AI responses across realistic mental health conversations.

BenchmarksOpen story →

How UK AISI and EvalEval Are Making Benchmark Results Reproducible

Inconsistent benchmark testing has made it hard to compare AI systems fairly, creating a need for standardized evaluation methods.

BenchmarksOpen story →

Vals, backed by Andreessen Horowitz, is looking to become the gold standard for AI benchmarking

Vals AI is hoping to make AI benchmarking a more neutral and trustworthy resource in a world increasingly inundated by AI models.

Relevant courses

Want to go deeper than the news? Explore live, cohort-based AI courses taught by practitioners.

Browse AI courses on Maven →
Get Agent5 in your inbox

The AI stories worth your attention, and a way to test your read. Free, no spam.