PAPER·May 4, 2026Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria ScoringHugging Face Papers
TECH·April 30, 2026Catching both math and coding... Apple unveils new technology to improve LLM reasoning accuracy디지털투데이
TECH·April 25, 2026Accenture and Wavemaker Partner...Address AI Code Hallucination Issues with a Two-Pass Architecture디지털투데이
PAPER·April 20, 2026QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading StrategiesHugging Face Papers
TECH·April 19, 2026Even the best AI models lose about half their performance when charts get complicated, new benchmark findsTHE DECODER
PAPER·April 6, 2026SQUIRE: Interactive UI Authoring via Slot Query Intermediate RepresentationsApple Machine Learning Research