L'AI la Meilleure? GPT-5.1 vs GPT-5 Pro vs Gemini 3?

L'AI la Meilleure? GPT-5.1 vs GPT-5 Pro vs Gemini 3?

Which AI is REALLY the Best? GPT-5.1 vs GPT-5 Pro vs Gemini?

🎙 Parlons IA 👥 17K 📅 November 13, 2025 ⏱ 16 min 👁 2K 📄 expert opinion 🧭 2026-09-08
Available in: English (current) Français

Keywords

GPT-5.1GPT-5 ProGeminiAI codingAI agents

Summary

The video presents a comparative analysis of several AI models, focusing on coding capabilities, agentic behavior, and conversational style. The creator tests GPT-5.1, GPT-5 Pro, and Gemini 2.5 on a Super Mario-style game coding task, noting that GPT-5 Pro delivers a fully functional game while GPT-5.1 produces a buggy and incomplete version. A second test on generating nebulae shows Gemini 2.5 outperforming GPT-5.1 in visual quality and menu integration. In an agentic task involving legal research, GPT-5.1 takes longer and deviates from instructions by including irrelevant blog sources, raising reliability concerns. The video also discusses new parameters for controlling reasoning and verbosity, and highlights a shift towards more human-like, conversational responses, which the creator views as a potential pitfall. The creator warns against over-reliance on AI for financial advice, citing examples of misleading influencer content. The video concludes by asking viewers for their preferences between the models.

147 words

Critical Evaluation

Value of the Information & Strength of the Argument

The video offers practical, hands-on comparisons that are valuable for developers and AI enthusiasts. The coding tests provide concrete examples of performance differences, and the agentic test highlights real-world reliability issues. However, the argumentation is largely anecdotal and lacks rigorous methodology. The creator’s personal preferences and subjective judgments (e.g., on writing style) are presented without objective criteria. The warning about AI-generated financial advice is pertinent, but the claim that ‘99% of content is false’ is an unsubstantiated generalization. Overall, the information is useful for initial impressions but not sufficient for definitive conclusions.

Scientific Rigor, Source Quality, Title Accuracy

The video does not cite specific scientific sources or studies, relying instead on personal tests and observations. The description includes links to the creator’s own platforms and promotional materials, but no external references. The title accurately reflects the content, which is a comparison of AI models. The lack of verifiable sources and the absence of peer-reviewed references reduce the scientific rigor. The creator’s methodology is not detailed, and the tests are not reproducible without the exact prompts and settings used. The video’s adéquation with its title is good, but the content is more of an opinion piece than a rigorous analysis.

208 words

Title / Content Match

The title accurately reflects the content, which compares GPT-5.1, GPT-5 Pro, and Gemini models.

Quality & Reliability

6/10

The video provides hands-on comparative tests of AI models, but lacks rigorous methodology, clear metrics, and verifiable sources. The creator's personal opinions and anecdotal evidence dominate, reducing overall reliability.

Chapters

Cited Sources

Concurring Sources

  • OpenAI official blog — Official OpenAI announcements about model updates and features.
  • Google AI blog — Google's AI blog, relevant for Gemini updates.

Dissenting Sources

  • No specific discordant sources found — The video does not reference any sources that contradict its claims; however, the lack of external validation means potential discrepancies with independent benchmarks are not addressed.

Contribution & Novelties

The video provides a timely, hands-on comparison of the latest AI models, highlighting practical differences in coding and agentic performance. It introduces new parameters for controlling reasoning and verbosity, which are useful for users. The discussion on the humanization of AI responses and its potential risks adds a valuable perspective.

Pour aller plus loin :

  • GPT-5.1 — Official OpenAI page for GPT-5.1, though specific details may be limited.
  • Gemini — Google’s Gemini platform, relevant for comparison.
  • Replika — AI companion app mentioned in the video, illustrating human-AI relationships.
  • Mixture of Experts — Technical concept underlying many modern AI models.
  • AI alignment — Discusses the challenge of ensuring AI follows human intentions, relevant to the agentic test.

116 words

Radar Profile

The radar profile shows moderate scores across all dimensions, with quantity of information slightly higher than quality and reliability. This suggests the video offers a decent amount of content but lacks depth and verifiability, making it more suitable for casual viewers than for those seeking rigorous analysis.

Reliability 5/10