Vivold Consulting
Other

Why Cohere’s ex-AI research lead is betting against the scaling race

Key Insights

Former Cohere AI research head Sara Hooker argues that the industry's obsession with ever-larger language models has hit diminishing returns. Her new venture focuses on efficient architectures and data quality, not sheer scale.

Stay Updated

Get the latest insights delivered to your inbox

Betting against the AI scaling race

Sara Hooker, once leading research at Cohere, has become one of the most vocal critics of the industry's 'bigger is always better' mentality. In her recent talk and interview with TechCrunch, she argues that model scaling has plateaued — and that breakthroughs will come from smarter training data, specialized models, and adaptive reasoning systems.

What Hooker is saying

  • Massive models like GPT-4, Claude 3, and Gemini Ultra show marginal accuracy gains but at exponential cost in compute and energy.
  • The new research focus is data curation — improving signal-to-noise ratios and dynamic sampling instead of blindly increasing corpus size.
  • Hooker calls for a future of modular AI — smaller, task-specific systems that communicate rather than one monolithic model.

The technical case


  • Studies from Hooker’s lab show that training efficiency doubles when data is pre-filtered for reasoning diversity rather than scale.

  • Mixture-of-experts architectures are regaining attention, offering large-model performance at a fraction of cost.

  • She advocates for open, interpretable benchmarks beyond MMLU and HELM to measure real-world reliability.

Why this matters for the AI ecosystem


  • The “scaling ceiling” could reshape the AI race. Firms like Anthropic and OpenAI may pivot toward data efficiency and reasoning-enhanced training.

  • VC interest is shifting from GPU-driven startups to optimization-focused ventures building inference-efficient stacks.

  • For enterprise buyers, the new competitive edge becomes: Can you get 90% of GPT-4 performance for 10% of the cost?

The human and policy angle


  • Hooker also highlights the carbon footprint of scaling and urges policymakers to include compute transparency standards.

  • “We’ve mistaken size for intelligence,” she says. “Now we have to make AI more human by making it smaller.”

More in Other

All Other stories

Coinbase for Agents: Automating portfolio trading with AI

Coinbase for Agents connects AI agents to live financial execution, letting them trade, pay, and rebalance within user-defined limits straight from a portfolio. It offers a CLI path for dev tools like Claude Code and Codex and a Model Context Protocol path for web agents like ChatGPT and Claude, with agents confined to isolated portfolios and run through Know-Your-Transaction checks. It completes a stack that began with AgentKit (2024) and the x402 agent-payments protocol, turning LLMs from advisors into actors.

Microsoft's open source tools were hacked to steal passwords of AI developers

Microsoft disabled dozens of its open-source GitHub projects - at least 70 - after hackers reportedly injected password-stealing malware into the code. Many affected projects relate to Azure and tools used with AI coding apps like Claude Code, the Gemini CLI, and VS Code, with credentials stolen when developers opened the compromised tools. It's reportedly Microsoft's second such breach in weeks, described as a re-compromise of a previously hit project.

Antigravity 2.0: a platform to orchestrate autonomous AI agents

Google expanded Antigravity, its agent-first development platform, beyond coding into a system for developing and managing cohorts of autonomous AI agents - headlined by Antigravity 2.0, a standalone desktop app that acts as a central home for orchestrating agents across tasks. It runs on a specially optimized version of Gemini 3.5 Flash that Google says is 12x faster than other frontier models. Users could start trying the experience at I/O.