From Directions to Regions: Decomposing Activations in Language Models via Local Geometry
从方向到区域:通过局部几何分解语言模型中的激活
Or Shafran, Shaked Ronen, Omri Fahn, Shauli Ravfogel, Atticus Geiger, Mor Geva
机构
*
Blavatnik School of Computer Science(Blavatnik计算机科学学院)
;
AI, Tel Aviv University, Israel(人工智能,特拉维夫大学,以色列)
;
New York University, New York, NY, USA(纽约大学,纽约,纽约州,美国)
Learning Markov Decision Processes under Fully Bandit Feedback
在完全老虎机反馈下学习马尔可夫决策过程
Zhengjia Zhuo, Anupam Gupta, Viswanath Nagarajan
机构
*
Department of Industrial and Operations Engineering, University of Michigan(工业与运作工程系,密歇根大学)
;
Computer Science Department, New York University(计算机科学系,纽约大学)
MedAraBench: Large-Scale Arabic Medical Question Answering Dataset and Benchmark
MedAraBench:大规模阿拉伯语医学问答数据集和基准测试
Mouath Abu-Daoud, Leen Kharouf, Omar El Hajj, Dana El Samad, Mariam Al-Omari, Jihad Mallat, Khaled Saleh, Nizar Habash, Farah E. Shamout
机构
*
Engineering Division, New York University Abu Dhabi(纽约大学阿布扎克分校工程系)
;
Cleveland Clinic Abu Dhabi(阿布扎克克利夫兰诊所)
;
Science Division, New York University Abu Dhabi(纽约大学阿布扎克分校科学系)
The Strategic Foresight of LLMs: Evidence from a Fully Prospective Venture Tournament
大语言模型的战略前瞻性:来自一个前瞻性创业竞赛的证据
Felipe A. Csaszar, Aticus Peterson, Daniel Wilde
机构
*
Ross School of Business University of Michigan(密歇根大学罗斯商学院)
;
NYU Stern School of Business New York University(纽约大学 Stern 商学院)
;
Kelley School of Business Indiana University(印第安纳大学凯利商学院)
Lotus: Efficient LLM Training by Randomized Low-Rank Gradient Projection with Adaptive Subspace Switching
Lotus: 通过随机低秩梯度投影与自适应子空间切换实现高效的LLM训练
Tianhao Miao, Zhongyuan Bao, Lejun Zhang
机构
*
Hong Kong Baptist University, School of Science, Hong Kong, China(香港 Baptist 大学,科学学院,香港,中国)
;
Fudan University, School of Data Science, Shanghai, China(复旦大学,数据科学学院,上海,中国)
;
New York University, Tandon School of Engineering, New York, USA(纽约大学,Tandon 工程学院,纽约,美国)
机构
*
Simons Institute for the Theory of Computing, UC Berkeley(Simons理论计算研究所,伯克利大学)
;
New York University(纽约大学)
;
Massachusetts Institute of Technology(麻省理工学院)
Semantic-Aware Advanced Persistent Threat Detection Using Autoencoders on LLM-Encoded System Logs
基于语义的高级持续性威胁检测:使用自动编码器对LLM编码的系统日志进行分析
Waleed Khan Mohammed, Zahirul Arief Irfan Bin Shahrul Anuar, Mousa Sufian Mousa Mitani, Hezerul Abdul Karim, Nouar AlDahoul
机构
*
Faculty of Artificial Intelligence and Engineering, Multimedia University Cyberjaya, Malaysia(人工智能与工程学院,多媒体大学(Cyberjaya校区))
;
Centre for Image and Vision Computing, Centre of Excellence for Artificial Intelligence, Faculty of Artificial Intelligence and Engineering, Multimedia University, Cyberjaya, Selangor, Malaysia(图像与视觉计算中心,人工智能卓越中心,人工智能与工程学院,多媒体大学(Cyberjaya校区,Selangor州))
;
Computer Science, New York University Abu Dhabi, UAE(计算机科学,纽约大学阿布扎克校区)
SurfelSoup: Learned Point Cloud Geometry Compression With a Probablistic SurfelTree Representation
SurfelSoup:基于概率Surfel树的端到端点云几何压缩框架
Tingyu Fan, Ran Gong, Yueyu Hu, Yao Wang
机构
*
Department of Electrical and Computer Engineering, New York University Tandon School of Engineering, Brooklyn, NY, USA(电气工程系,纽约大学塔能工程学院,布鲁克林,纽约,美国)
A Survey of AI Methods for Geometry Preparation and Mesh Generation in Engineering Simulation
工程仿真中几何准备和网格生成的AI方法综述
Steven Owen, Nathan Brown, Nikos Chrisochoides, Rao Garimella, Xianfeng Gu, Franck Ledoux, Na Lei, Roshan Quadros, Navamita Ray, Nicolas Winovich, Yongjie Jessica Zhang
机构
*
Sandia National Laboratories(桑迪亚国家实验室)
;
Old Dominion University(旧 Dominion 大学)
;
Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室)
;
New York University / Stony Brook University(纽约大学 / 斯通布鲁克大学)
;
CEA(法国原子能委员会)
;
Dalian University of Technology(大连理工大学)
;
Carnegie Mellon University(卡内基梅隆大学)