Anthropic and OpenAI are joining the AI stage at TechCrunch Disrupt 2026

At TechCrunch Disrupt 2026, the AI Stage is back to dig into the single hottest topic in the community for the past few years, presented by Google for Startups.

The top 3

  1. Highest MMLU Scores for General Knowledge: As of August 2026, OpenAI's GPT-5 leads the MMLU leaderboard with a score of 92.5%, followed by OpenAI's o1 at 91.8%, demonstrating superior general knowledge across 57 subjects.
  2. Top LLMs on Graduate-Level GPQA Benchmark: OpenAI's GPT-5.6 Sol and Anthropic's Claude Mythos Preview both lead the GPQA benchmark with 94.6% as of August 2026, showcasing their advanced reasoning in specialized scientific domains.
  3. Highest Scores in Agentic Coding & Reasoning: Claude Opus 5 achieves a leading 85.2% on SWE-Bench for agentic coding and 96.2% on GPQA Diamond for reasoning, while OpenAI's GPT-5.6 Sol leads overall on BenchLM's composite scores with 82 as of August 2026.

Sources

Open the full topic