PAPER·20 hours agoFewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adaptation in Coding AssistantsHugging Face Papers
PAPER·21 hours agoWould You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense ReasoningHugging Face Papers
PAPER·22 hours agoSULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based Surface Landmine Detection Under Domain ShiftHugging Face Papers
PAPER·yesterdayN_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile TokensHugging Face Papers