Robust Feature Learning for Multi-Index Models in High Dimensions
Comments 41 pages, 1 figure. To appear in the International Conference on Learning Representations (ICLR), 2025
期刊&会议
International Conference on Learning Representations · 会议 · Machine Learning
Comments 41 pages, 1 figure. To appear in the International Conference on Learning Representations (ICLR), 2025
Comments International Conference on Learning Representations
Comments Code and data at https://github.com/saprmarks/feature-circuits. Demonstration at https://feature-circuits.xyz
Journal ref International Conference on Learning Representations, 2025
Comments Published at ICLR 2025
Comments 8 pages, 2 figures, Accepted at the ICLR 2025 Workshop on Frontiers in Probabilistic Inference
Comments 36 pages, 2 figures. To appear in the International Conference on Learning Representations (ICLR), 2025
Comments Accepted to ICLR 2025 LMRL workshop (International Conference on Learning Representations, Learning Meaningful Representations of Life Workshop)
Comments Delta workshop at ICLR 2025
Comments ICLR 2025 ML4RS workshop
Comments ICLR 2025 Spotlight
Journal ref ICLR 2025
Comments Published as a workshop paper at SCOPE - ICLR 2025
Comments Camera ready. ICLR 2025
Comments International Conference on Learning Representations
Comments ICLR 2025. First three authors contributed equally. Code is available at https://github.com/thuml/DARE
Comments This paper is accepted by ICLR 2021 Robust and reliable machine learning in the real world Workshop
Comments Accepted to ICLR 2025
Journal ref ICLR 2025
Comments Accepted to International Conference on Learning Representations (ICLR) Workshop on Scalable Optimization for Efficient and Adaptive Foundation Models (SCOPE)
Journal ref In The Thirteenth International Conference on Learning Representations, 2025
Comments 13 pages, 9 figures
Journal ref ICLR 2025 Workshop on Building Trust in Language Models and Applications
Comments Accepted at ICLR 2025
Comments Accepted in ICLR 2025
Journal ref Addepalli, S., Varun, Y., Suggala, A., Shanmugam, K., & Jain, P. (2025). Does safety training of LLMs generalize to semantically related natural prompts? In The Thirteenth International Conference on Learning Representations 2025
Comments Accepted to the Thirteenth International Conference on Learning Representations (ICLR 2025)
Comments ICLR 2025 Camera-ready
Comments ICLR 2025
Comments Accepted by ICLR 2025
Comments Accepted by ICLR 2025 as Spotlight. Project homepage: https://interleave-eval.github.io/
Comments Accepted by ICLR2025 as spotlight paper; Project homepage: https://tum-pbs.github.io/ConFIG/
Journal ref The Thirteenth International Conference on Learning Representations, 2025
Comments Accepted by ICLR 2025
Comments Published as a conference paper at ICLR 2025