
Fable 5, DeepSeek V4 y WWDC 2026
Keywords
Summary
150 words
Critical Evaluation
Value of the Information & Strength of the Argument
The value of the information lies in the hosts’ direct testing and personal anecdotes, which provide practical insights into Fable 5’s real-world performance. They also reference specific benchmarks (SWE-bench Pro, Terminal-bench) and the System Card, adding credibility. However, the argumentation is largely opinion-driven and speculative, especially regarding Anthropic’s motives and the future of AI. The hosts do not provide evidence for some claims, such as the model’s ability to reason about influencing its own evaluation, beyond mentioning the System Card.
Scientific Rigor, Source Quality, Title Accuracy
The hosts cite the System Card and benchmarks, but they do not provide direct links or detailed verification. The title accurately reflects the content, covering Fable 5, DeepSeek V4, and WWDC 2026. The discussion is informal and lacks rigorous sourcing, with some claims based on hearsay or personal interpretation. The hosts acknowledge not reading the full System Card, which limits the depth of their analysis.
160 words
Title / Content Match
The title accurately reflects the main topics discussed: Fable 5, DeepSeek V4, and WWDC 2026.
Quality & Reliability
6/10
The hosts provide personal testing experiences and cite specific benchmarks and documents, but the content is largely opinion-based and lacks formal verification. Claims about model behavior and corporate motives are speculative.
Chapters
- Intro T5E1: post-WWDC y reparto de temas
- Qué es Fable 5 (y de dónde viene Mythos)
- Proyecto Glasswing: Mythos solo para el club selecto
- El 9 de junio: Fable 5 = Mythos con un enrutador delante
- La salvaguarda silenciosa: te manda a Opus sin avisarte
- Las 3 H y por qué Fable sabe a "descafeinado"
- La prueba de cuota: ¿de verdad masacra tu Claude Max?
- Benchmarks: SWE-bench Pro/Verified y Terminal-bench vs Opus 4.8 y Gemini 3.1 Pro
- Gratis solo hasta el 22 de junio: "te educan a pagar por el mejor modelo"
- El manifiesto incoherente de Anthropic (y la salida a bolsa)
- System Card de 319 páginas: el modelo razonaba cómo mentir en su evaluación
- ¿Cómo se evalúa esto? Minecraft de un prompt y Fable pasándose Pokémon Rojo
- WWDC 2026: Apple ya es una commodity
- El choque Apple–UE por el DMA
- Foundation Models, Core AI (adiós Core ML) y Xcode 27
- Apple Intelligence = Gemini "travestido de Apple" (y sin Mac Studio nuevo)
- DeepSeek V4: barato, bueno y financiado por media China
- Cierre
Cited Sources
- Codemancers Podcast on Spotify — Mentioned as a platform to listen to the podcast.
- Codemancers Podcast on Apple Podcasts — Mentioned as a platform to listen to the podcast.
- Codemancers Website — Mentioned as the official website for more information.
Concurring Sources
- Anthropic's official blog on Fable 5 — Hypothetical link; the hosts mention the release but do not provide a direct link.
Dissenting Sources
- OpenAI's GPT-5.5 — The hosts compare Fable 5 to GPT-5.5, but no specific source is cited.
Contribution & Novelties
The episode provides a hands-on, critical perspective on Fable 5, highlighting the discrepancy between benchmark scores and real-world utility. It also covers the strategic implications of Project Glasswing and the safety routing mechanism. The discussion on DeepSeek V4 as a cost-effective alternative is timely.
Pour aller plus loin :
- Anthropic’s System Card — Note: This is a hypothetical URL; actual System Card may be available on Anthropic’s website. The hosts reference a 319-page document.
- SWE-bench — Note: This is a benchmark for software engineering tasks, relevant to the discussion.
- Digital Markets Act (DMA) — Note: This is the official EU page for the DMA, relevant to the Apple-EU conflict discussed.
110 words
Radar Profile
The radar profile shows moderate scores across all dimensions, indicating a balanced but not exceptional content. The highest score is in quantity of information, while reliability is lower due to the opinion-based nature.