Publication InsightTok: Improving Text and Face Fidelity in Discrete Tokenization for Autoregressive Image Generation Yang Yue, Fangyun Wei, Tianyu He, Jinjing Zhao, Zanlin Ni, Zeyu Liu, Junliang Guo, Lei Shi, Yue Dong, Li Chen, Ji Li, Gao Huang, Dong Chen May 2026 arXiv | May 2026
Publication MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays Emre Hayir, Lorin Crawford, Alex Lu May 2026 arXiv | May 2026
Publication PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning Heegeon Yoon, Eunseop Yoon, Jinqiu Hong, Soo-Hwan Eom, Gwanhyeong Koo, M. Hasegawa-Johnson, Qi Dai, Chong Luo, C. D. Yoo CVPR 2026 | May 2026
Publication No One Knows the State of the Art in Geospatial Foundation Models Isaac Corley, Nils Lehmann, Caleb Robinson, Gabriel Tseng, Anthony Fuller, Hamed Alemohammad, Evan Shelhamer, Jennifer Marcus, Hannah Kerner May 2026 arXiv | May 2026
Publication Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models Andreas Bergmeister, Stefanie Jegelka, Nikolas Nusken, Carles Domingo-Enrich, Jakiw Pidstrigach May 2026 arXiv | May 2026
Publication From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation Bohan Li, Shuo Yang, Bao Peng, Xianda Guo, Erli Zhang, Youqi Tao, Junfeng Duan, Daguang Xu, Qi Dou, Xin Jin, Wenjun Zeng, Hao Zhao, Yueming Jin May 2026 arXiv | May 2026
Publication Unifying Scientific Communication: Fine-Grained Correspondence Across Scientific Media Megha Mariam K.M, Vineeth N Balasubramanian, C. V. Jawahar CVPR 2026 | May 2026
Publication Morphology prediction of small nanoparticles in any orientation from single electron micrographs Henrik Eliasson, Fangjinhua Wang, Xi (Ada) Wang, Dániel Baráth, Marc Pollefeys, Rolf Erni npj Computational Materials | May 2026
Publication Understanding Annotator Safety Policy with Interpretability Alexander X. Oesterling, Donghao Ren, Yannick Assogba, Dominik Moritz, Sunnie S. Y. Kim, Leon Gatys, Fred Hohman May 2026 arXiv | May 2026
Publication Audio-Visual Intelligence in Large Foundation Models Youxuan Qin, Kaihong Liu, Shengqiong Wu, Kai Wang, Shijian Deng, Yapeng Tian, Junbin Xiao, Yazhou Xing, Yinghao Ma, Bobo Li, Roger Zimmermann, Lei Cui, Furu Wei, Jiebo Luo, Hao Fei May 2026 arXiv | May 2026