Evaluating Fine-Tuned and Base Language Models in Maternal and Vaccination Healthcare for African SettingsResearchSep 22, 2026
An Empirical Cost Attribution of Context-Compression Gateways in Multi-Turn Coding AgentsResearchSep 22, 2026
Beyond Accuracy and Surface Fluency: Risk-Sensitive Evaluation of LLMs for Legal Clause GenerationResearchSep 22, 2026
GRRR: The Geometry of Reshaping, Rotation, and Routing in Decoder LLM post-trainingResearchSep 22, 2026
A Comparative Framework for Evaluating Foundation Models on Tabular Data: A Case Study in HealthcareResearchSep 22, 2026
Beyond the Text: Verifying That Agent-Written Papers Are Backed by Their ArtifactsResearchSep 22, 2026
Balancing Reasoning and Hardware Constraints in RAG Pipelines for Ukrainian Multi-Domain Document UnderstandingResearchSep 22, 2026
Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and ReasoningResearchSep 22, 2026