Vivold Consulting
Product & Tech Updates

Anthropic Unveils Claude Haiku 4.5: A Leap in Speed and Affordability

Key Insights

Anthropic has introduced Claude Haiku 4.5, its fastest and most cost-efficient AI model to date.

  • Performance: Matches Sonnet 4's capabilities in coding and agent tasks.

  • Benchmark: Achieves 73.3% on SWE-bench Verified, ranking among top coding models.

  • Availability: Accessible via Claude.ai on web, iOS, and Android platforms.

  • Pricing: Starts at $1 per million input tokens and $5 per million output tokens, with potential cost savings through prompt caching and batch processing.

Stay Updated

Get the latest insights delivered to your inbox

Why This Matters for Your AI Strategy

Anthropic's latest release, Claude Haiku 4.5, is reshaping the AI landscape by offering high performance at a fraction of the cost.

  • Unprecedented Speed and Efficiency:

  • Performance Parity: Matches the capabilities of the more robust Sonnet 4 in coding and agent tasks.

  • Benchmark Excellence: Scores 73.3% on SWE-bench Verified, placing it among the world's leading coding models.

  • Strategic Implications:

  • Cost-Effective Scaling: With pricing starting at $1 per million input tokens and $5 per million output tokens, businesses can achieve up to 90% cost savings through prompt caching and 50% via batch processing.

  • Versatile Deployment: Available on Claude.ai across web, iOS, and Android, facilitating seamless integration into existing workflows.
In essence, Claude Haiku 4.5 offers a compelling blend of speed, affordability, and performance, making it a strategic asset for enterprises aiming to scale their AI capabilities efficiently.

More in Product & Tech Updates

All Product & Tech Updates stories

Claude gets a 'Reflect' dashboard: Spotify Wrapped for your AI habit - and a masterclass in retention design

Anthropic launched Reflect, a beta dashboard (Free, Pro, and Max users with Memory on) that visualises Claude usage over 1-12 months - top topics, task types, peak hours - and coaches you via its 4D AI Fluency Framework (delegation, description, discernment, diligence), suggesting features like Projects or custom skills based on your patterns. It ships wellbeing controls (quiet hours, break nudges, reflection prompts) built with MIT Media Lab and Boston Children's Hospital experts, and excludes health-linked conversations entirely. TechCrunch's sharp read: beneath the mindfulness framing, Reflect showcases how much of your work runs through Claude - a retention play as much as a wellness one.

Gemini Spark lands on the Mac: Google's 24/7 agent starts working your local files

Gemini Spark, Google's 24/7 agentic assistant, is now available on Mac (beta, US-only, Google AI Ultra subscribers), where it can work directly with files on the computer - sorting and organising them, or turning a folder of invoices into a budgeting worksheet in Google Workspace. The update adds long-requested Google Tasks and Keep integrations plus third-party hooks into Canva, Dropbox, Instacart, OpenTable, and Zillow Rentals, real-time tracking of topics like stocks and breaking news, and - notably for builders - custom MCP support for wiring in your own apps. It puts Spark in direct competition with Claude Desktop, Microsoft Copilot, and OpenClaw for the desktop, where the real productivity (and governance) questions live.

L'Oreal brings Maybelline virtual try-on to ChatGPT

L'Oreal has announced a wide-ranging collaboration with OpenAI, unveiled at VivaTech 2026, that brings Maybelline's virtual makeup try-on directly into ChatGPT via L'Oreal's ModiFace AR technology. The deal spans consumer shopping tools, product discovery for brands like Lancome and Kerastase, advertising pilots (SkinCeuticals, CeraVe, Garnier), and R&D - including using OpenAI's GPT-Rosalind life-sciences model for skin-microbiome research. It lands as OpenAI reports ChatGPT at more than 900 million weekly users.