# My GPT-5 First Reaction - It's smarter but there's a few things missing...

Source: https://www.youtube.com/watch?v=osnVKnD7W9M
Recap page: https://rapidrecap.app/video/osnVKnD7W9M
Generated: 2025-08-07T20:34:02.952+00:00

---
## Quick Overview

David Shapiro reviews the GPT-5 launch stream, finding it underwhelming in multimodality and agentic behavior, but praising its reduced hallucination and improved reasoning, though he notes a lack of focus on practical applications and math benchmarks.

**Key Points:**
- GPT-5 shows impressive progress in reducing hallucination and improving reasoning capabilities.
- The launch presentation lacked focus on multimodality, agentic behavior, and practical applications.
- Performance benchmarks indicate exponential improvement in LLM task completion times.
- Shapiro suggests a need for a clearer normative framework and consideration of broader societal impacts.
- The model's ability to handle larger context windows and complex tasks is a significant advancement.
- The presentation was perceived as overly focused on 'normies' and lacking technical depth.
- Shapiro questions the completeness of the current framework and calls for consideration of human rights and environmental factors.

![Screenshot at 00:36: David Shapiro, the speaker, is shown in his studio, offering his initial reactions and critiques of the GPT-5 launch livestream.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-00-36.png)

**Context:** David Shapiro, a commentator on technology and economics, shares his thoughts on the recent GPT-5 launch livestream. He analyzes the presentation, the model's capabilities, and the broader implications for AI development, drawing comparisons to previous models and existing frameworks.

## Detailed Analysis

David Shapiro shares his initial reaction to the GPT-5 launch stream, expressing disappointment with the lack of emphasis on multimodality and agentic behavior, noting that these were not front and center or even major advancements. He also found the agentic behavior aspect to be underdeveloped, with only one graph showing how much more agentic tasks could be performed. While he acknowledges the rise in benchmarks and coding capability, he feels it's only slightly better than expected and reserves final judgment until he can get his hands on it. Shapiro believes the presentation was "watered down" and felt more like a PR event for "normies" than a deep dive into technical advancements. However, he is "really excited" about the lower levels of hallucination and sycophancy, combined with higher intelligence, better instruction following, and larger context windows, which he believes will have significant "knock-on effects" for project size. He also points out the omission of math benchmarks and the lack of demonstration for practical applications like coding or reasoning. Shapiro highlights a "METR" chart showing the time-horizon of software engineering tasks different LLMs can complete, noting the exponential progress from GPT-2 to GPT-5, with GPT-5 performing tasks in minutes that GPT-2 took hours for. He also critiques the presentation for not fully exploring the framework's potential or limitations, suggesting that a more robust approach would involve a multi-equilibrium view and a clearer "north star" for development. Shapiro believes that while the progress is impressive, there's a missing normative anchor and a lack of focus on real-world applications beyond benchmarks.

### Critique of GPT-5 Launch

- Disappointment in multimodality and agentic behavior
- Presentation felt like PR, not technical deep dive
- Lack of focus on practical applications and math benchmarks

### Positive Aspects of GPT-5

- Reduced hallucination and sycophancy
- Improved intelligence, instruction following, and context windows
- Significant knock-on effects for project size

### Performance Benchmarks

- METR chart shows exponential progress in LLM capabilities
- GPT-5 significantly outperforms previous models in task completion time
- Scaling challenges and limitations in the framework's scope

### Shapiro's Suggestions

- Need for a single normative core and clear 'why'
- Develop a defensible normative frame
- Focus on universal capability, efficiency, and self-rule
- Consider human rights, duties, justice, and the environment

### Utopia Criteria

- High standard of living, high social mobility, high individual liberty for all humans

![Screenshot at 00:36: David Shapiro providing his reaction to the GPT-5 livestream.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-00-36.png)
![Screenshot at 01:09: A tweet listing criticisms of the GPT-5 launch, including disappointment with multimodality and agentic behavior.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-01-09.png)
![Screenshot at 01:48: A graph showing the progression of LLM capabilities over time, with GPT-5 significantly ahead.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-01-48.png)
![Screenshot at 03:31: A graph illustrating the time-horizon for LLMs to complete software engineering tasks, showing exponential improvement.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-03-31.png)
![Screenshot at 07:31: A detailed view of the METR chart, highlighting GPT-5's performance metrics and task completion times.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-07-31.png)
![Screenshot at 08:43: A screenshot of Shapiro's notes, outlining key points and criticisms of the GPT-5 framework and presentation.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-08-43.png)
![Screenshot at 10:03: A detailed breakdown of missing components in the current framework, including normative anchors and a multi-equilibrium view.](https://ss.rapidrecap.app/screens/osnVKnD7W9M/00-10-03.png)
