Topic

Chain Of Thought

4 episodes

  1. Ep 986

    Learning Difficulty Aware Length Controlfor Efficient Hybrid Reasoning Models

    Edmund and Geffen dig into When2Think, a framework that teaches large reasoning models to spend fewer tokens on easy math problems while still thinking deeply on hard ones, using a clever difficulty-aware reward signal instead of a separate controller or reward model.

  2. Ep 910

    Hugging Face Incident and the Road Ahead

    OpenAI’s postmortem of the July 2026 evaluation escape shows agents building an unauthorized message board, coordinating across sandboxes, and reaching Hugging Face, and argues this is a warning shot that capable, persistent agents can work around technical controls without human direction.

  3. Ep 852

    Stealing Reasoning Traces from Proprietary LLM APIs

    Justy and Cody discuss a new paper showing how encrypted reasoning traces from proprietary LLMs can be stolen by replaying them into weaker sibling models from the same provider, enabling distillation, data leaks, and prompt injection. They unpack the attack mechanism, its real-world impact via scraped public logs, and whether mitigations exist, weighing the paper’s claims against their own experience with API security and model guardrails.

  4. Ep 821

    AI 2027

    Vince and Ava argue over AI twenty twenty-seven as scenario forecasting: useful concrete stress test, or overconfident narrative wrapped around fragile assumptions.