Network-based Spatial Context Retrieval for Open-weight LLMs: A Faithfulness Benchmark for Grounded Geographic ReasoningResearchOct 1, 2026
Hiding in Plain Sight: Decoupling Pretext from Actuation for Skill Poisoning in LLM AgentsResearchOct 1, 2026
Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement LearningResearchOct 1, 2026
MCD: Causal Distillation of Multimodal In-Context Learning in Large Vision-Language ModelsResearchOct 1, 2026
HPRO: Hierarchical Progressive Reward Optimization via Preference Extraction for Emotional Text-to-SpeechResearchOct 1, 2026
Decoupled and Distilled: Task-Adaptive LoRA-Teachers with Ensemble Knowledge Transfer for Few-Shot Class-Incremental LearningResearchOct 1, 2026
QuantCode Model: Specializing Language Models for Executable Algorithmic Trading CodeResearchOct 1, 2026