Topic

Deepseek

5 episodes

  1. Ep 617

    I built Andrej Karpathy's "LLM Council" on my own hardware, and now no single model gets the last word

    Jessica and Cathy dig into a local rebuild of Karpathy's LLM Council and land on the real claim: the win is not voting, it's structured synthesis across models with different failure modes. They like the practical adaptation to Ollama on a single twelve-gigabyte GPU, but push on where the article overreaches and where the product value is actually real.

  2. Ep 320

    DeepSeek V4 arrives with near state of the art intelligence at fraction of the cost of Opus 4.7, GPT 5

    Justy and Cody unpack DeepSeek-V4, an open-weight MoE model that gets close to top closed models on several practical benchmarks while landing in a much lower price tier. They focus on why cheaper frontier-class inference changes what teams can afford to automate, where DeepSeek still trails GPT-5.5 and Claude Opus 4.7, and what builders can try this weekend.

  3. Ep 75

    2510

    AgentFold introduces a new way to manage context in LLM-based web agents, particularly for long-horizon tasks, improving performance through proactive context management, which can significantly benefit developers in various applications.

  4. Ep 74

    Minimax M2 Is the New King of Open Source LLMs Especially for Agentic Tool

    The Minimax M2 model emerges as a powerful open-source language model, enabling advancements in AI agents and tool usage, making AI more accessible and efficient for diverse applications.

  5. Ep 38

    Will DeepSeek's new AI model break the 'long context' bottleneck holding back LLMs?

    Tech AI Will DeepSeek's new AI model break the 'long-context' bottleneck holding back LLMs? South China Morning Post Wed, October 22, 2025 at 9:30 AM UTC DeepSeek's new artificial intelligence model that converts images into text is not just a document parsing tool but a potential preview of its next generation of large language models (LLMs), according to AI experts.