TECH·6 days agoDeconstructing the Kimi K3 Technical Report: Moonshot AI Starts Setting Problems for DeepSeek as Two Scaling Axes Redefine the FrontierPandaily
TECH·6 days agoGemini 3.5 Pro Delays Explained and Why Google Shifts to Gemini 3.6 FlashGeeky Gadgets
TECH·July 25, 2026Anthropic's Claude Opus 5 delivers near-Fable 5 performance at half the token priceTHE DECODER
TECH·July 25, 2026Claude Opus 5 is here, and Anthropic says it can rival Fable 5 in some tasksDigital Trends
TECH·July 22, 2026Poolside Releases Laguna S 2.1, an Open-Weight Agentic Coding Model Punching Above Its Weight Class on SWE-Bench MultilingualMarkTechPost
TECH·June 3, 2026Microsoft's MAI-Code-1-Flash Scores 51% on SWE-Bench Pro with Just 5B Active Params | Hacker NewsHacker News
PAPER·June 1, 2026GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language ModelsHugging Face Papers
TECH·May 29, 2026Claude Opus 4.8 vs ChatGPT 5.5: A Stepping Stone to Anthropic’s Mythos SeriesGeeky Gadgets
TECH·May 28, 2026DeepSWE AI Coding Model Benchmark Finally Solves AI Training Data ContaminationGeeky Gadgets