Claude Sonnet 4.5 Ranked Safest LLM from Open-Source Audit Tool Petri

October 7, 2025

Key Insights

Claude Sonnet 4.5 emerges as the top-performing model in 'risky tasks,' surpassing GPT-5 in early evaluations by Petri, Anthropic's new open-source AI auditing tool. ([infoq.com](https://www.infoq.com/anthropic/news/?utm_source=openai))

Stay Updated

Get the latest insights delivered to your inbox

Is Your AI Strategy Prioritizing Safety?

In the latest evaluations by Petri, an open-source AI auditing tool developed by Anthropic, Claude Sonnet 4.5 has outperformed competitors, including GPT-5, in handling 'risky tasks.' This positions it as a leading model in AI safety.

Why This Matters:

- Enhanced Trust: Superior performance in risky scenarios builds confidence in AI applications.

- Competitive Edge: Adopting safer AI models can differentiate your offerings in the market.

- Regulatory Compliance: Aligning with top safety standards may ease compliance with emerging AI regulations.

Considering Claude Sonnet 4.5 for your AI initiatives could bolster both safety and competitiveness. ([infoq.com](https://www.infoq.com/anthropic/news/?utm_source=openai))

Source: infoq.com

Related Articles

Salesforce Unveils AI-Powered Slack Makeover with 30 New Features

Salesforce has announced a major update to Slack, introducing over 30 new AI-driven features aimed at enhancing workplace productivity and collaboration. Key enhancements include: - Advanced Slackbot capabilities for drafting content, summarizing conversations, and answering queries. - Integration with Salesforce CRM and third-party apps to provide context-aware assistance. - Proactive recommendations during video calls, such as surfacing relevant Salesforce records when key names are mentioned.

Salesforce Ramps Up Agentic AI Research with New Foundry Project

Salesforce has launched the AI Foundry, a new initiative aimed at accelerating agentic AI research and development. The project focuses on: - Bridging foundational research and product innovation through collaboration with strategic customers and academic partners. - Developing AI tools for high-impact enterprise areas, including simulated environments for testing AI agents and enhancing solutions like Agentforce Voice. - Exploring ambient intelligence to provide proactive, context-aware assistance without constant user input.

VHA Deploys Salesforce-Powered Agentic Operating System, Saving Thousands of Staff Hours for Front-Line Veteran Care

The Veterans Health Administration (VHA) has implemented a Salesforce-powered agentic operating system, resulting in significant operational efficiencies. Key outcomes include: - Transitioning from static reporting to automated problem-solving, eliminating administrative silos. - Freeing thousands of staff hours, allowing more focus on direct Veteran support. - Creating a connected performance management layer, enhancing care delivery across facilities.