LLM intelligence hub

Track the frontier of LLMs without losing the evidence.

EvalKit brings public leaderboards, model profiles, benchmark guides, and curated LLM news into one citation-first workspace.

Leaderboard rows464
Model profiles410
Public sources8
Verified claims0

Current leaders

Start with the models people are already comparing.

Open full explorer →

LLM news

Curated releases, research, and benchmark shifts.

Read news →

AI was supposed to win people over by now — it hasn’t

As AI becomes harder to avoid, consumers are growing more wary of the technology — and Silicon Valley is discovering that widespread adoption doesn’t necessarily lead to acceptance.

TechCrunch AISource

Amazon makes its AI-powered Alexa+ free on Fire TV, no Prime required

Amazon is making its AI-powered Alexa+ assistant free on all compatible Fire TV devices in the U.S., automatically upgrading users whether or not they subscribe to Prime.

TechCrunch AISource

Cognition CEO denies report that SpaceX tried to acquire the startup

SpaceX was reportedly in talks to buy AI coding startup Cognition. SpaceX has already acquired Cursor as it races to catch up to rivals like OpenAI and Anthropic in enterprise AI.

TechCrunch AISource

Google Gemini is getting a dedicated student hub

As we're gearing up for back-to-school season, Google is rolling out a new dedicated student hub in Gemini. It's a one-stop repository for collecting research in a study notebook, creating flashcards, taking practice quizzes, and more. Goo...

The Verge AISource

Benchmark guide

Read scores like a product decision, not a scoreboard.

Open guide →

Trust policy

No fake “tested by us” claims.

Public rows are labeled as replicated public-source data or editorial context. “Verified by EvalKit” stays at 0 until there is real run evidence attached.

Read citation policy