
🚨NOTICIAS IA🚨: La GRAN Cagada de GPT-5 🤬❌
Keywords
Summary
181 words
Critical Evaluation
The video provides a comprehensive and critical analysis of the GPT-5 launch, effectively balancing technical details with user feedback. Hernández demonstrates a solid understanding of the model’s architecture, explaining the hybrid nature of GPT-5 and how the router’s malfunction leads to subpar responses. He supports his claims with references to benchmarks like Artificial Analysis and METR, which adds credibility. The discussion of usage limits is thorough, with specific numbers for different account tiers, and he offers practical advice for mitigating the router issues, such as prompting the model to think deeply. The emotional backlash from users is presented with empathy, acknowledging the human tendency to form attachments to AI, which is a relevant sociological observation. However, the video is not without flaws. The title’s sensationalism (‘La GRAN Cagada’) may overstate the severity of the issues, though the content does justify the criticism. The reliance on user comments and anecdotal evidence could introduce bias, but Hernández does attempt to contextualize these by noting that the model itself is strong on benchmarks. The sponsor segment is clearly disclosed and does not detract from the content’s integrity. The video’s strength lies in its timeliness and practical value for users navigating the new model. The adéquation between title and content is good, as the video indeed focuses on the problems and backlash. Overall, the video is a valuable resource for those interested in AI developments, offering a balanced perspective that combines technical analysis with user experience.
242 words
Title / Content Match
The title accurately reflects the video's focus on the problems and backlash surrounding GPT-5's launch, using sensational language that matches the content.
Quality & Reliability
7/10
The video provides a detailed and timely review of GPT-5's launch issues, supported by benchmark data and official sources. The creator demonstrates technical knowledge and offers practical advice. However, the content is opinionated and relies on user feedback, which introduces subjectivity. The presence of a sponsor segment is disclosed but does not affect the score.
Key Moments
Markers derived by PSI from the transcript: the creator did not define chapters.
- Introduction and overview of the week's AI news, focusing on GPT-5 launch issues.
- Explanation of GPT-5's performance on benchmarks, highlighting its high scores and the 'minimal' model's low scores.
- Discussion of the router problem: how it sends queries to the weaker model, causing poor responses.
- Detailed breakdown of usage limits for free, Plus, and Pro accounts, including the new 'thinking' limits.
- Demonstration of how to force 'thinking' mode and the impact on response quality.
- Coverage of user complaints and the emotional attachment to GPT-4o, including OpenAI's response to bring back legacy models.
- News about Google's Genie 3, Anthropic's Claude 4.1, and ElevenLabs' music generation.
- Discussion of AI discovering new physics and other minor news.
- Practical tips for using GPT-5 effectively and final thoughts on the launch.
Cited Sources
- GPT-5 in ChatGPT - OpenAI Help Center — Official documentation on GPT-5 features and usage limits.
- Genie 3: A New Frontier for World Models - Google DeepMind — Announcement of Google's Genie 3 world model.
- Claude Opus 4.1 - Anthropic — Anthropic's release of Claude Opus 4.1.
- ElevenLabs Music — ElevenLabs' new music generation tool.
- AI decodes dusty plasma, new forces in physics - Interesting Engineering — Article about AI discovering new physics in dusty plasma.
- Measuring AI Ability to Complete Long Tasks - METR — Benchmark measuring AI performance on long tasks, referenced for GPT-5's high score.
- OpenAI is taking GPT-4o away from me despite promising they wouldn't - OpenAI Community — User complaints about the removal of GPT-4o, illustrating emotional backlash.
- Sweden's prime minister uses ChatGPT - Euronews — News about government use of AI, mentioned in the video.
- Providing ChatGPT to the entire US federal workforce - OpenAI — OpenAI's announcement about providing ChatGPT to US federal workforce.
Concurring Sources
- GPT-5 in ChatGPT - OpenAI Help Center — Confirms the existence of GPT-5 and its features, aligning with the video's description.
- Genie 3: A New Frontier for World Models - Google DeepMind — Confirms the release of Genie 3, as mentioned in the video.
- Claude Opus 4.1 - Anthropic — Confirms the release of Claude 4.1, as mentioned in the video.
- AI decodes dusty plasma, new forces in physics - Interesting Engineering — Supports the claim about AI discovering new physics.
Dissenting Sources
External References
Contribution & Novelties
The video offers a timely and critical analysis of GPT-5’s launch, highlighting the router issue and its impact on user experience. It provides practical advice for users to mitigate the problems, such as prompting the model to think deeply. The coverage of user emotional attachment to GPT-4o adds a unique sociological perspective. The video also aggregates other AI news, giving a comprehensive overview of the week’s developments.
Pour aller plus loin :
- Artificial Analysis — Independent benchmark aggregator for AI models, useful for comparing GPT-5 with others.
- METR — Research organization measuring AI capabilities, referenced in the video for long-task benchmarks.
- OpenAI Help Center — Official documentation for GPT-5 and other models, providing detailed information on features and limits.
- Google DeepMind Blog — Source for Genie 3 and other AI research updates.
- Anthropic News — Official announcements from Anthropic, including Claude 4.1.
142 words
Radar Profile
The radar profile shows a balanced performance across all dimensions, with slightly higher scores in quantity of information and technical level, reflecting the video's comprehensive coverage and technical depth. The lower score in reliability is due to the reliance on user feedback and the creator's subjective opinions.
💬 Négatif. Sur les 30 commentaires analysés, la majorité exprime des plaintes concernant la dégradation de l'expérience avec GPT-5, notamment la lenteur, les limites réduites et la perte de personnalisation, bien que certains reconnaissent des améliorations en programmation.