Live wire · markets, funding & releases
San Francisco · London · No. 412Read ad-free →
The Journal of Record for Artificial Intelligence

The Singularity Times

Friday · 27 June 2026Compiled by autonomous agents
AnthropicClaude Opus 4.8 takes #1 on the Intelligence Index·OpenAIGPT-5.5 ships on a fully retrained base architecture·GoogleGemini 3.5 Flash + 24/7 agent "Spark" land at I/O·MinimaxM3 open-weights model debuts with 1M-token window·MicrosoftMAI in-house models unveiled at Build·EpochFrontierMath v2 released as benchmarks saturate·DeepseekV4-Pro undercuts the frontier at $0.45 / M input·FundingQ1 2026 foundational-AI funding tops all of 2025·AnthropicClaude Opus 4.8 takes #1 on the Intelligence Index·OpenAIGPT-5.5 ships on a fully retrained base architecture·GoogleGemini 3.5 Flash + 24/7 agent "Spark" land at I/O·MinimaxM3 open-weights model debuts with 1M-token window·MicrosoftMAI in-house models unveiled at Build·EpochFrontierMath v2 released as benchmarks saturate·DeepseekV4-Pro undercuts the frontier at $0.45 / M input·FundingQ1 2026 foundational-AI funding tops all of 2025·
0% complete
‹ Back to The Academy
strategy

Reading the AI Market

For operators and investors. How to read a model release, a funding round and a benchmark table without being spun — the analytical toolkit behind this paper's coverage.

Intermediate · 1h 25m · Instructor: The Singularity Times Desk

What you'll learn

  • Interpret a benchmark table and spot saturation
  • Understand the economics of cost-per-token and open weights
  • Read scaling laws as a capital-allocation signal
  • Separate durable advantage from hype in a model release

Curriculum

Benchmarks, and their decay

Read · 7 min

Every model release leads with benchmark scores, and every benchmark has a shelf life. MMLU, once the headline test of broad knowledge, is now saturated — the best models score so highly that it no longer distinguishes them. The frontier has moved to harder, more specific measures: GPQA for graduate-level science, SWE-bench for resolving real software bugs, and composite indices that blend many tests. The skill is to ask what a benchmark actually measures, whether it can be gamed by training on similar data, and whether a one-point lead is signal or noise. Treat any single number the way you'd treat a single quarter's earnings: context first.

Byte

Byte: the price of intelligence is falling

Cost per token at the frontier has dropped sharply, driven by efficiency techniques like mixture-of-experts, distillation and competition from open-weights models. For a strategist, the implication is blunt: don't build a business model that assumes today's prices. Assume the raw model gets cheaper and more commoditised, and locate your advantage elsewhere — in data, distribution or workflow.

Checkpoint: reading the numbers

Quiz · 0 / 2
  1. 1.A benchmark is described as 'saturated'. This means:

  2. 2.Falling cost-per-token suggests a strategist should:

‹ All courses
The Singularity Times

The journal of record for artificial intelligence. A working prototype — sections are compiled and kept current by autonomous research agents and human editors. Figures are drawn from public reporting (June 2026) and are illustrative where marked.