Grok 3 vs ChatGPT : je fais le test COMPLET !

Grok 3 vs ChatGPT : je fais le test COMPLET !

🎙 Ludo Salenne 👥 267K 📅 February 18, 2025 ⏱ 43 min 👁 51K 📄 news review 🧭 2026-08-21
Available in: English (current) Français

Keywords

Grok 3ChatGPTcomparisonAI testDeep Research

Summary

In this video, Ludo Salenne tests Grok 3, the new AI from Elon Musk’s xAI, and compares it with ChatGPT. He starts by explaining how to access Grok 3, either through X or the dedicated website with a VPN. He then conducts a series of tests: a basic question about the universe, a request for an unfiltered opinion on Emmanuel Macron, and an attempt to generate an image of Macron, which succeeds impressively. He also tests image generation of himself, showing the ability to iteratively modify images. The video then compares Grok 3’s DeepSearch feature with ChatGPT’s Deep Research on the topic of caffeine’s effects on mental health, noting that Grok 3 completed the task in 1 minute 21 seconds with 80 sources, while ChatGPT took 7 minutes with 29 sources. He also tests Grok 3’s ability to handle a PDF with a trap, compares it with DeepSeek, tests vision capabilities, and evaluates its integration with X. The video concludes with pricing information and the creator’s overall opinion, highlighting Grok 3’s strengths in speed and uncensored responses, but noting some limitations in reasoning and content moderation.

186 words

Critical Evaluation

Value of the Information & Strength of the Argument

The video provides valuable hands-on demonstrations of Grok 3’s capabilities, including image generation, DeepSearch, and integration with X. The creator’s argumentation is based on direct testing and comparison, which adds credibility. However, the analysis is largely subjective, and the creator does not delve into technical specifications or benchmark scores beyond mentioning a leaderboard. The tests are practical and cover a range of use cases, but the reasoning for why Grok 3 is ’the most intelligent’ is not thoroughly substantiated.

Scientific Rigor, Source Quality, Title Accuracy

The video cites the LMArena leaderboard as evidence of Grok 3’s performance, but does not provide a direct link. The creator mentions that ChatGPT’s Deep Research uses the O3 model, which is a valid point. The description includes links to the creator’s own resources and other tutorials, but no external scientific sources. The title accurately reflects the content, and the video is well-structured with clear chapters. The creator’s methodology is transparent, but the lack of external references limits the scientific rigor.

175 words

Title / Content Match

The title accurately reflects the content: a comprehensive test of Grok 3 compared to ChatGPT, covering various features and use cases.

Quality & Reliability

7/10

The video is a hands-on comparative test of Grok 3 against ChatGPT, with clear methodology and multiple test scenarios. The creator demonstrates practical usage and provides subjective but informed opinions. However, the analysis lacks depth in technical details and relies on personal impressions rather than standardized benchmarks.

Chapters

Cited Sources

Concurring Sources

  • LMArena Leaderboard — The video mentions this leaderboard as evidence of Grok 3's top ranking.

Dissenting Sources

  • OpenAI's Deep Research — The video compares Grok 3's DeepSearch with ChatGPT's Deep Research, but does not provide a direct link. This source is provided for context.

Contribution & Novelties

The video provides a timely and practical comparison of Grok 3 with ChatGPT, offering hands-on demonstrations of features like image generation and DeepSearch. It highlights Grok 3’s speed and uncensored nature, which are notable differentiators. The creator’s approach of testing with real-world prompts adds practical value for viewers.

Pour aller plus loin :

  • LMArena Leaderboard — Official leaderboard for AI models, referenced in the video.
  • Deep Research in ChatGPT — Official OpenAI page explaining Deep Research, which is compared in the video.
  • Grok 3 by xAI — Official page for Grok, providing details on the model and its features.
  • Caffeine and Mental Health — A scientific review on caffeine’s effects on mental health, relevant to the test topic.

118 words

Radar Profile

The radar profile shows high scores in quantity of information and fiabilite, reflecting the video's comprehensive coverage and practical testing. The lower score in niveau technique indicates that the content is accessible but lacks deep technical analysis.

Reliability 7/10

💬 Très positif. Sur les 30 commentaires analysés, les spectateurs expriment une forte appréciation pour la vidéo, saluant la qualité des tests et la réactivité du créateur, avec quelques demandes de vidéos supplémentaires sur le classement des IA.