GPT 5.5 Full Review & 5 Mind-Blowing Use Cases

GPT 5.5 Full Review & 5 Mind-Blowing Use Cases

🎙 Pat Simmons 👥 24K 📅 April 24, 2026 ⏱ 13 min 👁 2K 📄 expert opinion 🧭 2026-09-07
Available in: English (current) Français

Keywords

GPT-5.5Codexbenchmarksuse casesAI capabilities

Summary

Pat Simmons reviews OpenAI’s GPT-5.5, released in April 2026, by running it through a series of practical tests. The video begins with a benchmark comparison, noting a 13-point lead over Claude’s Opus. The first test involves building a 3D model of Monica’s apartment from Friends using Three.js in Codex; the result is functional but visually inaccurate, earning a B-. Next, he replicates an Artemis II space mission app from OpenAI’s docs, which closely matches the demo and uses real NASA data, earning an A-. He then builds a dungeon game and a UFO tank shooter, both working in one shot, with the dungeon game requiring a second prompt to fix visibility and controls. A quick SVG test of a pelican on a bicycle produces a result similar to Simon Willison’s. The video then shifts to knowledge work: a financial modeling task generates a comprehensive 3-year projection Excel file with multiple scenarios and a self-check tab in 7 minutes, which the creator finds impressive. Writing tests include a salary negotiation email, a text to a friend about declining a bridesmaid role, and a Shakespearean sonnet about a Wi-Fi router; all are deemed effective. Finally, a reasoning lightning round covers classic gotchas like the number of R’s in ‘strawberry’ and the 9.11 vs 9.9 comparison, all answered correctly. The creator concludes that GPT-5.5 is a significant step up from 5.4, especially in coding and reasoning, and plans to use it in daily workflows.

241 words

Critical Evaluation

Value of the Information & Strength of the Argument

The video’s value lies in its practical, hands-on demonstrations of GPT-5.5’s capabilities across diverse tasks, from coding to writing and reasoning. The creator provides specific prompts and shows the outputs, allowing viewers to assess the model’s performance. The argumentation is based on anecdotal evidence from a single user’s tests, which is not statistically robust but offers a realistic glimpse into the model’s strengths and weaknesses. The creator acknowledges limitations, such as the apartment build’s inaccuracy, and provides a balanced perspective. However, the lack of a controlled methodology and the absence of comparison with other models in the same conditions weaken the overall argument.

Scientific Rigor, Source Quality, Title Accuracy

The video is a personal review, not a scientific study. The creator cites OpenAI’s official documentation and a benchmark from Simon Willison, but does not provide links to these sources in the description. The description only includes links to the creator’s own bootcamp and newsletter, which are promotional. The title accurately reflects the content, and the video’s structure is clear. The creator’s claims about GPT-5.5’s performance are not independently verified, and the video’s reliance on subjective impressions limits its scientific rigor. The creator does not discuss potential biases or limitations of his testing methodology.

212 words

Title / Content Match

The title accurately reflects the content: a full review of GPT-5.5 with multiple use cases demonstrated.

Quality & Reliability

6/10

The video is a hands-on review of GPT-5.5, based on the creator's own tests and demonstrations. It lacks formal methodology, peer review, or independent verification, and the creator's expertise is not established. However, the tests are reproducible and the claims are specific, lending some credibility.

Chapters

Cited Sources

  • AI Bootcamp — Promotional link for the creator's AI bootcamp, mentioned in the video description.
  • AI For Mortals Newsletter — Link to the creator's newsletter, mentioned in the video description.

Concurring Sources

  • OpenAI's GPT-5.5 announcement — The video's claims about GPT-5.5's benchmark performance align with OpenAI's official statements.

Dissenting Sources

  • Independent benchmarks — The video's anecdotal tests may not reflect the model's performance in other contexts, and independent benchmarks could yield different results.

Contribution & Novelties

The video provides a practical, hands-on evaluation of GPT-5.5, demonstrating its capabilities in real-world tasks such as 3D modeling, game development, financial modeling, and creative writing. It offers a user’s perspective on the model’s strengths and weaknesses, complementing official benchmark data. The creator’s tests are reproducible, and he shares his prompts, adding value for viewers interested in replicating the results.

Pour aller plus loin :

  • OpenAI’s GPT-5.5 announcement — Official information about the model’s capabilities and benchmarks.
  • Three.js documentation — Reference for the 3D library used in the coding tests.
  • Simon Willison’s pelican SVG benchmark — Context for the SVG test.
  • Artemis II mission — Official NASA page for the mission used in the space app test.

117 words

Radar Profile

The radar profile shows a balanced performance across the four dimensions, with slightly higher scores in information quantity and technical level, reflecting the video's hands-on demonstrations. The lower reliability score indicates the subjective nature of the review.

Reliability 5/10