PAPER·yesterdayCoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action ModelarXiv
PAPER·yesterdayReasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reasoning ModelsHugging Face Papers
PAPER·yesterdayFrom Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic SearchHugging Face Papers
HARDWARE·3 days agoA Modder’s RTX 4080 Was Enough To Play AAA Games, But Not For Running LLMs, So He Integrated NVIDIA’s Tesla V100 At A Throwaway Price To Run 27B AI ModelsWccftech
TECH·July 20, 2026Alibaba Previews Qwen3.8-Max, a 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open-Weight LaunchMarkTechPost
CODING·July 14, 2026Systems Engineering Playbook: Optimizing Qwen 3.5-397B MoE on Ironwood (TPU7x)Google Developers Blog