| 2025 | |
|---|---|
![]() | PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding Jang Hyun Cho, Andrea Madotto, Effrosyni Mavroudi, Triantafyllos Afouras, Tushar Nagarajan, Muhammad Maaz, Yale Song, Tengyu Ma, Shuming Hu, Suyog Jain, Miguel Martin, Huiyu Wang, Hanoona Rasheed, Peize Sun, Po-Yao Huang, Daniel Bolya, Nikhila Ravi, Shashank Jain, Tammy Stark, Shane Moon, Babak Damavandi, Vivian Lee, Andrew Westbury, Salman Khan, Philipp Krähenbühl, Piotr Dollár, Lorenzo Torresani, Kristen Grauman, Christoph Feichtenhofer NeurIPS 2025 code arxiv |
![]() | Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation Xiuyu Yang, Shuhan Tan, Philipp Krähenbühl ICCV 2025 code arxiv project |
![]() | Robust Autonomy Emerges from Self-Play Marco Cusumano-Towner, David Hafner, Alex Hertzberg, Brody Huval, Aleksei Petrenko, Eugene Vinitsky, Erik Wijmans, Taylor Killian, Stuart Bowers, Ozan Sener, Philipp Krähenbühl, Vladlen Koltun ICML 2025 arxiv |
![]() | Cut Your Losses in Large-Vocabulary Language Models Erik Wijmans, Brody Huval, Alexander Hertzberg, Vladlen Koltun, Philipp Krähenbühl ICLR 2025 code arxiv |
![]() | Distilling Structural Representations into Protein Sequence Models Jeffrey Ouyang-Zhang, Chengyue Gong, Yue Zhao, Philipp Krähenbühl, Adam R Klivans, Daniel J Diaz ICLR 2025 code arxiv |
![]() | Does Spatial Cognition Emerge in Frontier Models? Santhosh Kumar Ramakrishnan, Erik Wijmans, Philipp Kraehenbuehl, Vladlen Koltun ICLR 2025 code arxiv |
![]() | Image and Video Tokenization with Binary Spherical Quantization Yue Zhao, Yuanjun Xiong, Philipp Krähenbühl ICLR 2025 code arxiv |
![]() | Language-Image Models with 3D Understanding Jang Hyun Cho, Boris Ivanovic, Yulong Cao, Edward Schmerling, Yue Wang, Xinshuo Weng, Boyi Li, Yurong You, Philipp Krähenbühl, Yan Wang, Marco Pavone ICLR 2025 arxiv project |







