Formal & Physical Sciences AI — Models & ResearchENOn Policy Distillation - Using LLMs to train better LLMsNeural Breakdown with AVBAugust 23, 2026 30 min★ ★ ★ ★ ☆ 4/5On-Policy DistillationKnowledge DistillationLLM Training
Formal & Physical Sciences AI — Models & ResearchENI trained a Reasoning Language Model with RL on an unverifiable taskNeural Breakdown with AVBAugust 16, 2026 36 min★ ★ ★ ★ ☆ 4/5Reinforcement LearningLanguage ModelsReward Modeling
Formal & Physical Sciences AI — Models & ResearchENSmall Language Model Alignment - Finetune SLMs to ALWAYS pick the best answer (Unsloth DPO)Neural Breakdown with AVBMay 30, 2026 34 min★ ★ ★ ★ ☆ 4/5DPOPreference OptimizationSmall Language Models
Formal & Physical Sciences AI — Models & ResearchENDeepSeek V4 so powerful, but how is it so CHEAP? (A deep dive into Sparse Attention)Neural Breakdown with AVBApril 30, 2026 20 min★ ★ ★ ★ ☆ 4/5DeepSeek V4Sparse AttentionKV Cache
Applied Sciences & Engineering AI — Models & ResearchENTiny Language Models - How to build INSANELY FAST local models! (Unsloth, Outlines)Neural Breakdown with AVBApril 18, 2026 46 min★ ★ ★ ★ ☆ 4/5Fine-TuningSynthetic DataConstrained Decoding
Formal & Physical Sciences AI — Models & ResearchENHow to finetune LLMs on custom data domains (CPT tutorial with Unsloth)Neural Breakdown with AVBMarch 29, 2026 24 min★ ★ ★ ★ ☆ 4/5Fine-TuningContinued Pre-TrainingLoRA
Formal & Physical Sciences AI — Models & ResearchENRecursive Language Models (RLMs) - Let's build the coolest agents ever! (Theory & Code)Neural Breakdown with AVBFebruary 21, 2026 49 min★ ★ ★ ★ ☆ 4/5Recursive Language ModelsLLMInference
Formal & Physical Sciences AI — Models & ResearchENLet's train Vision Language Models (VLM) from scratch using just Text-Only LLMs!Neural Breakdown with AVBJanuary 30, 2026 30 min★ ★ ★ ★ ☆ 4/5Vision Language ModelQ-FormerBLIP-2
Formal & Physical Sciences AI — Models & ResearchENDeepSeekV4 - Manifold Constrained Hyper Connections (mHC) and the evolution of ResNetsNeural Breakdown with AVBJanuary 11, 2026 22 min★ ★ ★ ★ ☆ 4/5Deep LearningResidual ConnectionsHyper-Connections
Formal & Physical Sciences AI — Models & ResearchENHow to train Multi Agent Collaborative Agents with Reinforcement Learning (CTDE Explained)Neural Breakdown with AVBDecember 4, 2025 21 min★ ★ ★ ★ ☆ 4/5Reinforcement LearningMulti-AgentPPO
Applied Sciences & Engineering AI — Models & ResearchENHow to build your own long-term Agentic Memory System for LLMs | Mem0 from scratch in DSPyNeural Breakdown with AVBOctober 23, 2025 51 min★ ★ ★ ★ ☆ 4/5LLMMemoryDSPy
Formal & Physical Sciences AI — Models & ResearchENA visual guide on Reinforcement Learning - the 6 things that makes it “click”Neural Breakdown with AVBSeptember 14, 2025 33 min★ ★ ★ ★ ☆ 4/5Reinforcement LearningMachine LearningQ-Learning
Formal & Physical Sciences AI — Models & ResearchENLet me explain PyTorch in 7 ConceptsNeural Breakdown with AVBAugust 16, 2025 42 min★ ★ ★ ★ ☆ 4/5PyTorchDeep LearningNeural Networks