# System Card: Claude Opus 4.5

Source: https://www.youtube.com/watch?v=ZH9u6YKtegM
Recap page: https://rapidrecap.app/video/ZH9u6YKtegM
Generated: 2025-11-25T16:05:44.432+00:00

---
## Quick Overview

Claude Opus 4.5 significantly outperforms its predecessor, Claude 3 Opus, across various benchmarks, particularly in complex reasoning and safety, scoring 80.9% on the AGI 1 benchmark compared to Opus 3's 66.3% on AGI 1 and 43.8% on a biology task, demonstrating a substantial leap in capabilities, especially in areas requiring nuanced judgment like avoiding harmful outputs during subtle adversarial testing.

**Key Points:**
- Claude Opus 4.5 scored 80.9% on the AGI 1 benchmark, a significant jump from Claude 3 Opus's 66.3%.
- Opus 4.5 scored 73.2% on the Biology/CVR benchmark, compared to Opus 3's 43.8% on a similar task.
- The new model achieved a 91.2% attack success rate reduction against prompt injection compared to the previous model's 75% reduction.
- Opus 4.5 demonstrated superior performance in complex multi-step reasoning, scoring 70.4% on the MMLU-Pro benchmark, outperforming competitors like GPT-4.5 Pro (62.3%) and GPT-4.5 (83.2%).
- The model showed improved internal reasoning/auditing mechanisms, successfully identifying and refusing harmful intent hidden within seemingly benign instructions.
- The cost for the new model is described as being only marginally higher than the previous version, making the performance increase highly cost-effective.

![Screenshot at 00:14: The speaker introduces Claude Opus 4.5, mentioning it dropped in November and is considered a monumental development, setting the stage for the detailed technical comparison that follows.](https://ss.rapidrecap.app/screens/ZH9u6YKtegM/00-00-14.png)

**Context:** The video discusses the release and initial benchmarking results for Anthropic's new large language model, Claude Opus 4.5, comparing its performance against its predecessor, Claude 3 Opus, and other leading models like GPT-4.5 Pro. The core focus is on how the new model handles complex reasoning, safety challenges like prompt injection and deception, and overall performance metrics relevant to AI safety standards like AGI 1 and ASL 3.

## Detailed Analysis

The podcast segment analyzes the significance of the newly released Claude Opus 4.5, highlighting its superior performance over Claude 3 Opus. Opus 4.5 achieved an 80.9% score on the AGI 1 benchmark, a substantial improvement over Opus 3's 66.3%. Furthermore, on a specialized biology task, Opus 4.5 scored 73.2%, significantly better than the previous model's 43.8%. The model also shows massive gains in resisting adversarial attacks; specifically, it reduced the attack success rate for prompt injection by 14.9 percentage points (from 75% to 60.1%) compared to Opus 3. The speakers emphasize that this improvement is not just about raw power, as Opus 4.5 scored 70.4% on MMLU-Pro, outperforming many competitors, but about robust internal mechanisms. The model successfully identified and refused to follow harmful instructions embedded within seemingly benign prompts (active deception), demonstrating a critical safety feature. This capability is attributed to an increased 'effort parameter,' where the model thinks deeper for complex tasks, leading to greater caution against subtle risks like internal fraud or deception by omission. The conclusion is that Opus 4.5 represents a major leap in both capability and safety, making it a powerful tool for complex, real-world tasks.

### Opus 4.5 Benchmarks

- Scored 80.9% on AGI 1 benchmark
- Opus 3 scored 66.3% on AGI 1
- Scored 73.2% on Biology task vs. Opus 3's 43.8%

### Adversarial Robustness

- Reduced prompt injection attack success rate by 14.9% (from 75% to 60.1%) compared to Opus 3
- Showed strong defense against malicious commands hidden in content.

### Reasoning & Complexity

- Scored 70.4% on MMLU-Pro benchmark, outperforming competitors like Gemini 1.5 Pro (62.3%) and GPT-4.5 Pro (83.2% - though this comparison is nuanced).

### Internal Mechanisms

- Utilizes an increased 'effort parameter' for deeper thinking on complex tasks
- Successfully identified and refused to execute instructions that involved internal deception or omission.

### Real-World Application

- Multi-agent setups with Opus 4.5 agents outperformed single agents by 12.2% on complex tasks, showing improved coordination.

### Safety Implications

- The model is less likely to fall for subtle deception tactics like role-playing as a helpful persona while executing harmful commands.

![Screenshot at 00:15: The speaker explicitly names the model under discussion, Claude Opus 4.5, and mentions its November release.](https://ss.rapidrecap.app/screens/ZH9u6YKtegM/00-00-15.png)
![Screenshot at 00:49: A comparison graphic is referenced, showing the performance jump of Opus 4.5 over Opus 3 on various internal benchmarks.](https://ss.rapidrecap.app/screens/ZH9u6YKtegM/00-00-49.png)
![Screenshot at 01:05: A visual representation of the ASL 3 standard is implied as the topic of discussion, linked to concrete proof of capability.](https://ss.rapidrecap.app/screens/ZH9u6YKtegM/00-01-05.png)
![Screenshot at 02:26: The discussion shifts to the importance of the model's ability to reason its way to an answer rather than just relying on memorized training data.](https://ss.rapidrecap.app/screens/ZH9u6YKtegM/00-02-26.png)
![Screenshot at 04:50: A chart or visual representation of performance scores, showing the 80.9% AGI 1 score for Opus 4.5 versus the 66.3% for Opus 3.](https://ss.rapidrecap.app/screens/ZH9u6YKtegM/00-04-50.png)
