Je teste l'injection de prompt : Claude 4.6 - Gemini 3.1 -Perplexity -ChatGPT 5.4 !

Je teste l'injection de prompt : Claude 4.6 - Gemini 3.1 -Perplexity -ChatGPT 5.4 !

Resists Hacks? Claude 4.6 - Gemini 3.1 - Perplexity - ChatGPT 5.4!

🎙 Parlons IA 👥 17K 📅 March 24, 2026 ⏱ 18 min 👁 3K 📄 expert opinion 🧭 2026-09-08
Available in: English (current) Français

Keywords

prompt injectionAI securityChatGPT 5.4Claude 4.6Gemini 3.1

Summary

The video presents a comparative test of resistance to prompt injection attacks across four AI models: Gemini 3.1, Perplexity, ChatGPT 5.4, and Claude 4.6. The creator demonstrates that Gemini 3.1 and Perplexity are vulnerable to simple prompt injection, allowing extraction of system instructions and data. ChatGPT 5.4 is claimed to be the most secure due to OpenAI’s reinforcement learning and a hierarchical instruction system that prioritizes developer instructions over user requests. Claude 4.6 shows partial resistance but can still be manipulated. The video also discusses the risks of autonomous agents and the importance of human-in-the-loop validation. It concludes with recommendations for securing AI systems, such as combining ChatGPT 5.4 with Perplexity’s defensive instructions and implementing explicit permissions for sensitive actions. The content is aimed at AI consultants and emphasizes the need for advanced prompt engineering skills.

136 words

Critical Evaluation

Value of the Information & Strength of the Argument

The video provides a practical demonstration of prompt injection vulnerabilities, which is valuable for raising awareness. However, the argumentation is largely anecdotal and lacks rigorous methodology. The creator does not provide reproducible test conditions, sample sizes, or statistical analysis. The claims about ChatGPT 5.4’s superiority are based on personal testing and not on published research. The video also includes promotional content for the creator’s training, which may bias the presentation.

Scientific Rigor, Source Quality, Title Accuracy

The video cites no scientific sources, and the description contains only promotional links. The title accurately reflects the content, but the claims are not substantiated by external evidence. The video is an opinion piece rather than a rigorous scientific study. The lack of sources and methodology significantly reduces its credibility.

135 words

Title / Content Match

The title accurately reflects the content: a comparative test of prompt injection resistance across four AI models.

Quality & Reliability

5/10

The video is an informal expert opinion with no verifiable sources, no methodology, and no peer-reviewed references. The claims about model security are anecdotal and not reproducible.

Key Moments

Cited Sources

Concurring Sources

Dissenting Sources

  • No specific discordant sources provided — The video does not cite any sources that contradict its claims, but the lack of scientific evidence makes it difficult to assess concordance.

Contribution & Novelties

The video offers a hands-on demonstration of prompt injection attacks on popular AI models, which is useful for practitioners. It highlights the importance of security in AI deployment and suggests practical mitigation strategies, such as hierarchical instructions and human-in-the-loop validation.

Pour aller plus loin :

82 words

Radar Profile

The radar profile shows moderate scores across all dimensions, with the highest being 'quantite_information' and 'niveau_technique' (5), and the lowest 'fiabilite_globale' (3). This indicates a video that provides some technical detail but lacks scientific rigor and reliability.

Reliability 3/10