Publication Only Say What You Know: Calibration-Aware Generation for Long-Form Factuality Wen Luo, Guangyue Peng, Liang Wang, Nan Yang, Wei Li, Yuhan Song, Shaohang Wei, Feifan Song, Furu Wei, Houfeng Wang May 2026 arXiv | May 2026
Publication Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy Shakeel A. Sheikh, Patrick Marmaroli, Md. Sahidullah, Slim Ouni, Fabrice Hirsch, Goncalo Leal, Bjorn W. Schuller May 2026 arXiv | May 2026
Publication VeriTrail: Closed-Domain Hallucination Detection with Traceability Dasha Metropolitansky, Jonathan Larson 2026 ICLR’26 | April 2026 Video
Publication LLMs Corrupt Your Documents When You Delegate Philippe Laban, Tobias Schnabel, Jennifer Neville 2026 COLM | April 2026
Publication Pushing the Limits of On-Device Streaming ASR: A Compact, High-Accuracy English Model for Low-Latency Inference Nenad Banfic, David Fan, Kunal Vaishnavi, Sam Kemp, Sunghoon Choi, Rui Ren, Sayan Shaw, Meng Tang April 2026 arXiv | April 2026
Publication Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context Adaptation Zichong Li, Chen Liang, Liliang Ren, Tuo Zhao, Yelong Shen, Weizhu Chen April 2026 arXiv | April 2026
Publication Do Transformers Use their Depth Adaptively? Evidence from a Relational Reasoning Task A. Curth, Rachel Lawrence, Sushrut Karmalkar, Niranjani Prasad April 2026 arXiv | April 2026
Publication Discourse Diversity in Multi-Turn Empathic Dialogue Hongli Zhan, Emma S. Gueorguieva, Javier Hernandez, Jina Suh, Desmond C. Ong, Junyi Jessy Li April 2026 arXiv | April 2026 Project
Publication Evaluating Cooperation in LLM Social Groups through Elected Leadership Ryan Faulkner, Anushka Deshpande, David Guzman Piedrahita, Joel Z. Leibo, Zhijing Jin April 2026 arXiv | April 2026
Publication Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models Avni Mittal, Shanu Kumar, Sandipan Dandapat, Monojit Choudhury April 2026 arXiv | April 2026