Verification of PETSc with CIVL using LLM-generated ACSL contracts and deterministic driver generationResearchSep 29, 2026
DriveHierarchy: A Benchmark for Diagnosing VLM Driving Capabilities from Open-Loop Understanding to Closed-Loop ExecutionResearchSep 29, 2026
FIDAL: Diversity-Aware Federated Active Learning Under Real-World Distribution ShiftsResearchSep 29, 2026
Language-Augmented Video Action Anticipation: Design Fundamentals, Benchmarks, and Open ChallengesResearchSep 29, 2026
Replay in the Silent Degrees of Freedom: Continual Learning Without an Offline PhaseResearchSep 29, 2026
Enhancing generalization in endwall film cooling prediction: Incorporating the superposition principle into transformer-based neural operatorsResearchSep 29, 2026
ConflictVLA-Bench: Benchmarking Behavioral Responses of Vision-Language-Action Models to Premise ConflictsResearchSep 29, 2026
What Next-Event Accuracy Cannot See: Closed-Loop Evaluation of Emergency Department Trajectory SimulatorsResearchSep 29, 2026