SPACE: Semantic Projection and Alignment of CLIP Embeddings for Domain AdaptationResearchSep 22, 2026
Resist, Update, Reject: Preference Optimization Installs a Prior-Dependent Reliability SwitchResearchSep 22, 2026
DIPLOMAT: Dialogue-Span-Aware Direct Preference Optimization for Polite Persuasive Workplace Negotiation DialoguesResearchSep 22, 2026
Near-Optimal Primal-Dual Algorithm for Learning Linear Mixture CMDPs with Adversarial RewardsResearchSep 21, 2026
Elastic Threshold Attention: Learned Contextual Sparsity for Long-Context DecodingResearchSep 21, 2026
Continuous Delayed-Memory Stochastic Gradient Descent and Continuous-Time Reinforcement Learning from History of Astrophysical Time Series StudiesResearchSep 21, 2026
HERMES: Contrast-Aware Knowledge Graph Reasoning from Clinical Notes for Patient Outcome PredictionResearchSep 21, 2026
From Discharge Notes to Patient Understanding: Persona-Grounded, Open-Ended Simulation of LLMs as Discharge EducatorsResearchSep 21, 2026