Better Models, Faster Training: Sigmoid Attention for single-cell Foundation Models
Vijay Sadashivaiah, George Dasoulas, Judith Mueller, Soumya Ghosh
Action editor: Ole Winther

Better Models, Faster Training: Sigmoid Attention for single-cell Foundation Models
Vijay Sadashivaiah, George Dasoulas, Judith Mueller, Soumya Ghosh
Action editor: Ole Winther
On the Role of MLP Layers in Transformer ICL with Categorical Outcomes
Soumya Banerjee, Aaron T Wang, William Convertino et al.
Action editor: Alexander S. Ecker
The Intrinsic Dimension of Prompts in Internal Representations of Large Language Models
Karthik Viswanathan, Yuri Gardinazzi, Giada Panerai, Alberto Cazzaniga, Matteo Biagetti
Action editor: Zachary Charles