AI was supposed to win people over by now — it hasn’t
As AI becomes harder to avoid, consumers are growing more wary of the technology — and Silicon Valley is discovering that widespread adoption doesn’t necessarily lead to acceptance.
LLM intelligence hub
EvalKit brings public leaderboards, model profiles, benchmark guides, and curated LLM news into one citation-first workspace.
Current leaders
Reasoning leader: 72.51. leads on reasoning
LLM StatsGoogleCheapest top 10: $0.5/M. lowest input price
LLM StatsxAILongest context: 2.0M tokens. largest context window
LLM StatsAnthropicCoding leader: 57.83. wins on coding
LLM StatsMistral AIFastest output: 678.15 c/s. highest throughput
LLM StatsMoonshot AIBest open weights: 58.13. top open-weight score
LLM StatsLLM news
As AI becomes harder to avoid, consumers are growing more wary of the technology — and Silicon Valley is discovering that widespread adoption doesn’t necessarily lead to acceptance.
Amazon is making its AI-powered Alexa+ assistant free on all compatible Fire TV devices in the U.S., automatically upgrading users whether or not they subscribe to Prime.
SpaceX was reportedly in talks to buy AI coding startup Cognition. SpaceX has already acquired Cursor as it races to catch up to rivals like OpenAI and Anthropic in enterprise AI.
As we're gearing up for back-to-school season, Google is rolling out a new dedicated student hub in Gemini. It's a one-stop repository for collecting research in a study notebook, creating flashcards, taking practice quizzes, and more. Goo...
Benchmark guide
Trust policy
Public rows are labeled as replicated public-source data or editorial context. “Verified by EvalKit” stays at 0 until there is real run evidence attached.