Defusing Explosive Prompts: Understanding and Preventing Trigger-Based Prompt Injections in LLM AgentsResearchSep 22, 2026
DIPLOMAT: Dialogue-Span-Aware Direct Preference Optimization for Polite Persuasive Workplace Negotiation DialoguesResearchSep 22, 2026
Information-Geometric First-Passage Monitoring of Distributional Stability in Stochastic SystemsResearchSep 22, 2026
Paragraph Boundaries Are Not White Space:Compression Depth as the Signature of Hierarchical StructureResearchSep 22, 2026
Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful KnowledgeResearchSep 22, 2026
Deep Persona: A Psychologically Grounded Architecture and Evaluation Framework for Role-Playing Agents and SimulationsResearchSep 22, 2026