Signals

Filtered to Industry analysis · clear filter

Browse by topic

The frontier leaderboard is now a list of effort settings, not models

Of the top 15 entries on Artificial Analysis' intelligence index, most are the same handful of models at different reasoning-effort settings — Claude Opus 5 appears at max, xhigh, high and medium; GPT-5.6 Sol does the same. The spread within a single model is wide: Opus 5 ranges from 60.7 down to 56.3 across its…

Source ↗
AI engineering practiceIndustry analysisreasoning-effortbenchmarksmodel-selectionartificial-analysis

"Who's Afraid of Chinese Models?" — the open-weights gap becomes a strategy question

Simon Willison's latest lands alongside a wave of similar takes this week (Gary Marcus, Stratechery, and a 1,200+ point HN thread) arguing that China's open-weight labs — DeepSeek, Qwen, Kimi — have closed most of the capability gap to closed US models. Several independent voices converging on the same read in one…

Source ↗
Open-weight modelsIndustry analysischinaopen-weight-modelsdeepseekqwenindustry-analysis