Exploring Next / Topics / Neural Architecture Search Topic Neural Architecture Search 1 episode Ep 593 Jul 3, 2026 Morphing into Hybrid Attention Models Talon and Wildflower discuss FlashMorph, a new method for choosing which Transformer layers should keep full attention when converting pretrained LLMs into hybrid attention models. InferenceTrainingQwenFlashmorph