arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

International Conference on Learning Representations · 会议 · Machine Learning

共收录 9454
2410.12952 2025-03-04 cs.CL

Facilitating Multi-turn Function Calling for LLMs via Compositional Instruction Tuning

Mingyang Chen, Haoze Sun, Tianpeng Li, Fan Yang, Hao Liang, Keer Lu, Bin Cui, Wentao Zhang, Zenan Zhou, Weipeng Chen

Comments Accepted to ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12085 2025-03-04 cs.CR cs.AI cs.LG stat.ML

Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning

Fengyu Gao, Ruida Zhou, Tianhao Wang, Cong Shen, Jing Yang

Comments Accepted to ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11817 2025-03-04 cs.CV cs.LG cs.MM

Improving Long-Text Alignment for Text-to-Image Diffusion Models

Luping Liu, Chao Du, Tianyu Pang, Zehan Wang, Chongxuan Li, Dong Xu

Journal ref International Conference on Learning Representations (ICLR 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10781 2025-03-04 cs.CL cs.AI cs.LG

When Attention Sink Emerges in Language Models: An Empirical View

Xiangming Gu, Tianyu Pang, Chao Du, Qian Liu, Fengzhuo Zhang, Cunxiao Du, Ye Wang, Min Lin

Comments ICLR 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09400 2025-03-04 cs.CV

CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation

Yifeng Xu, Zhenliang He, Shiguang Shan, Xilin Chen

Comments ICLR 2025. Code: https://github.com/xyfJASON/ctrlora

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08190 2025-03-04 cs.CV cs.CR cs.GR cs.LG

Poison-splat: Computation Cost Attack on 3D Gaussian Splatting

Jiahao Lu, Yifan Zhang, Qiuhong Shen, Xinchao Wang, Shuicheng Yan

Comments Accepted by ICLR 2025 as a spotlight paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07672 2025-03-04 cs.CL cs.AI

MACPO: Weak-to-Strong Alignment via Multi-Agent Contrastive Preference Optimization

Yougang Lyu, Lingyong Yan, Zihan Wang, Dawei Yin, Pengjie Ren, Maarten de Rijke, Zhaochun Ren

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07295 2025-03-04 cs.SE cs.LG cs.PL

IterGen: Iterative Semantic-aware Structured LLM Generation with Backtracking

Shubham Ugare, Rohan Gumaste, Tarun Suresh, Gagandeep Singh, Sasa Misailovic

Comments Accepted at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07168 2025-03-04 cs.CL cs.SD eess.AS

Sylber: Syllabic Embedding Representation of Speech from Raw Audio

Cheol Jun Cho, Nicholas Lee, Akshat Gupta, Dhruv Agarwal, Ethan Chen, Alan W Black, Gopala K. Anumanchipalli

Comments Accepted at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07137 2025-03-04 cs.CL cs.AI cs.CR cs.LG

Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates

Xiaosen Zheng, Tianyu Pang, Chao Du, Qian Liu, Jing Jiang, Min Lin

Comments ICLR 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.06262 2025-03-04 cs.LG stat.ML

SymDiff: Equivariant Diffusion via Stochastic Symmetrisation

Leo Zhang, Kianoosh Ashouritaklimi, Yee Whye Teh, Rob Cornish

Comments Camera-ready version for ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05864 2025-03-04 cs.CL cs.AI

From Tokens to Words: On the Inner Lexicon of LLMs

Guy Kaplan, Matanel Oren, Yuval Reif, Roy Schwartz

Comments Accepted to the International Conference on Learning Representations (ICLR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05643 2025-03-04 cs.CV

TRACE: Temporal Grounding Video LLM via Causal Event Modeling

Yongxin Guo, Jingyu Liu, Mingda Li, Qingbin Liu, Xi Chen, Xiaoying Tang

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04870 2025-03-04 cs.LG stat.ML

On the Optimization and Generalization of Two-layer Transformers with Sign Gradient Descent

Bingrui Li, Wei Huang, Andi Han, Zhanpeng Zhou, Taiji Suzuki, Jun Zhu, Jianfei Chen

Comments 79 pages, 19 figures, ICLR 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04642 2025-03-04 cs.LG stat.ML

The Optimization Landscape of SGD Across the Feature Learning Strength

Alexander Atanasov, Alexandru Meterez, James B. Simon, Cengiz Pehlevan

Comments ICLR 2025 Final Copy, 40 Pages, 45 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04343 2025-03-04 cs.CL

Inference Scaling for Long-Context Retrieval Augmented Generation

Zhenrui Yue, Honglei Zhuang, Aijun Bai, Kai Hui, Rolf Jagerman, Hansi Zeng, Zhen Qin, Dong Wang, Xuanhui Wang, Michael Bendersky

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03601 2025-03-04 cs.LG cs.NA math.NA stat.ML

How Discrete and Continuous Diffusion Meet: Comprehensive Analysis of Discrete Diffusion Models via a Stochastic Integral Framework

Yinuo Ren, Haoxuan Chen, Grant M. Rotskoff, Lexing Ying

Comments Published at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03524 2025-03-04 cs.CL

Steering Large Language Models between Code Execution and Textual Reasoning

Yongchao Chen, Harsh Jhamtani, Srinagesh Sharma, Chuchu Fan, Chi Wang

Comments 32 pages, 12 figures, 12 tables

Journal ref The Thirteenth International Conference on Learning Representations (ICLR'2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03355 2025-03-04 cs.CV cs.AI

LANTERN: Accelerating Visual Autoregressive Models with Relaxed Speculative Decoding

Doohyuk Jang, Sihwan Park, June Yong Yang, Yeonsung Jung, Jihun Yun, Souvik Kundu, Sung-Yub Kim, Eunho Yang

Comments 30 pages, 13 figures, Accepted to ICLR 2025 (poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03115 2025-03-04 cs.CL

X-ALMA: Plug & Play Modules and Adaptive Rejection for Quality Translation at Scale

Haoran Xu, Kenton Murray, Philipp Koehn, Hieu Hoang, Akiko Eriguchi, Huda Khayrallah

Comments Published as a conference paper at ICLR 2025 (spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03011 2025-03-04 stat.ML cs.LG

Towards Understanding the Universality of Transformers for Next-Token Prediction

Michael E. Sander, Gabriel Peyré

Comments ICLR 2025, 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02392 2025-03-04 cs.LG math.AT

MANTRA: The Manifold Triangulations Assemblage

Rubén Ballester, Ernst Röell, Daniel Bīn Schmid, Mathieu Alain, Sergio Escalera, Carles Casacuberta, Bastian Rieck

Comments Accepted at ICLR 2025 (https://openreview.net/forum?id=X6y5CC44HM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02388 2025-03-04 cs.GT

Boosting Perturbed Gradient Ascent for Last-Iterate Convergence in Games

Kenshi Abe, Mitsuki Sakamoto, Kaito Ariu, Atsushi Iwasaki

Comments Accepted at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02268 2025-03-04 cs.LG cs.AI cs.CL cs.CV

Structural-Entropy-Based Sample Selection for Efficient and Effective Learning

Tianchi Xie, Jiangning Zhu, Guozu Ma, Minzhi Lin, Wei Chen, Weikai Yang, Shixia Liu

Comments Published as a conference paper at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02242 2025-03-04 cs.LG cs.AI

Robust Weight Initialization for Tanh Neural Networks with Fixed Point Analysis

Hyunwoo Lee, Hayoung Choi, Hyunju Kim

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01417 2025-03-04 cs.CV cs.AI cs.CL cs.LG

The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs

Hong Li, Nanxi Li, Yuanjie Chen, Jianbin Zhu, Qinlu Guo, Cewu Lu, Yong-Lu Li

Comments Accepted by ICLR 2025. Project page: https://mvig-rhos.com/llm_inception

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11219 2025-03-04 cs.CV cs.LG

Score Forgetting Distillation: A Swift, Data-Free Method for Machine Unlearning in Diffusion Models

Tianqi Chen, Shujian Zhang, Mingyuan Zhou

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.04591 2025-03-04 cs.CV cs.AI

HiLo: A Learning Framework for Generalized Category Discovery Robust to Domain Shifts

Hongjun Wang, Sagar Vaze, Kai Han

Comments v2: Accepted as a conference paper at ICLR 2025; Project page: https://github.com/Visual-AI/hilo/

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15589 2025-03-04 cs.CV cs.LG

Exploring the Effectiveness of Object-Centric Representations in Visual Question Answering: Comparative Insights with Foundation Models

Amir Mohammad Karimi Mamaghan, Samuele Papa, Karl Henrik Johansson, Stefan Bauer, Andrea Dittadi

Comments Published at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14985 2025-03-04 cs.CL cs.AI cs.LG

Generalization v.s. Memorization: Tracing Language Models' Capabilities Back to Pretraining Data

Xinyi Wang, Antonis Antoniades, Yanai Elazar, Alfonso Amayuelas, Alon Albalak, Kexun Zhang, William Yang Wang

Comments Accepted to ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏