36 results on this page · clear filters

machine learning Sep 15 Disentangling Representation Evolution in Transformers through Directional Decomposition by Shwai He, Haichao Zhang, Shen Yan machine learning Sep 15 ResSafe: Learning Safety Filtering with Residual Reinforcement Learning for Humanoids by Gechen Qu, Tong Zhang, Bike Zhang, Yen-Jen Wang, Koushil Sreenath, Claire Tomlin, Jason Jangho Choi machine learning Sep 15 A Chosen Future Can Still Be Rewritten: Causal Writability in Video Models by Xingyun Wang, Haomin Zheng, Man Yuan, Leqian Yang, Ziming Liu machine learning Sep 14 Type Diversity Enables Transformers to Generalise Compositionally by Anssi Moisio, Mathias Creutz, Mikko Kurimo machine learning Sep 14 MAxBench: A Multinomial Concept Recovery Benchmark by Divya Appapogu, Freya Behrens, Yonatan Belinkov, Aaron Mueller machine learning Sep 14 CanvasAnneal: Curriculum Reinforcement Learning for Diffusion Language Models by Blake Olson, Yuhang Song, Emmett McQuinn, Yuan Shangguan machine learning Sep 11 General Quantification of Covariate and Concept Shifts by Hongbo Chen, Li Charlie Xia machine learning Sep 11 Data Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data by Atindra Jha, Margaret Li, Jure Leskovec, Percy Liang, Luke Zettlemoyer machine learning Sep 11 Distance generalization in transformers: why bother with positional encoding? by Daniel Henrik Nevermann, Claudius Gros machine learning Sep 11 From Protocols to Evidence: Bounded Claims for AI in Service of the Common Good by Nitesh V. Chawla, Paulo Benanti