Samira Abnar
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
10
Venues
7
Active years
2016–2025
Best venue rank
A*
Where they publish
Papers
10 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICML | Parameters vs FLOPs: Scaling Laws for Optimal Sparsity for Mixture-of-Experts Language Models. | Samira Abnar, Harshay Shah, Dan Busbridge, Alaaeldin El-Nouby, Joshua M. Susskind, Vimal Thilak |
| 2023 | EMNLP | Scaling Laws vs Model Architectures: How does Inductive Bias Influence Scaling? | Yi Tay, Mostafa Dehghani, Samira Abnar, Hyung Won Chung, William Fedus, Jinfeng Rao, Sharan Narang, Vinh Q. Tran, Dani Yogatama, Donald Metzler |
| 2023 | ICLR | Diffusion Probabilistic Fields. | Peiye Zhuang, Samira Abnar, Jiatao Gu, Alexander G. Schwing, Joshua M. Susskind, Miguel ngel Bautista |
| 2022 | ICLR | Exploring the Limits of Large Scale Pre-training. | Samira Abnar, Mostafa Dehghani, Behnam Neyshabur, Hanie Sedghi |
| 2022 | ICLR | Scale Efficiently: Insights from Pretraining and Finetuning Transformers. | Yi Tay, Mostafa Dehghani, Jinfeng Rao, William Fedus, Samira Abnar, Hyung Won Chung, Sharan Narang, Dani Yogatama, Ashish Vaswani, Donald Metzler |
| 2021 | ICLR | Long Range Arena : A Benchmark for Efficient Transformers. | Yi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen, Dara Bahri, Philip Pham, Jinfeng Rao, Liu Yang, Sebastian Ruder, Donald Metzler |
| 2020 | AAAI | A Comparison of Architectures and Pretraining Methods for Contextualized Multilingual Word Embeddings. | Niels van der Heijden, Samira Abnar, Ekaterina Shutova |
| 2020 | ACL | Quantifying Attention Flow in Transformers. | Samira Abnar, Willem H. Zuidema |
| 2019 | CICLING | Robust Evaluation of Language-Brain Encoding Experiments. | Lisa Beinborn, Samira Abnar, Rochelle Choenni |
| 2016 | CIKM | The Healing Power of Poison: Helpful Non-relevant Documents in Feedback. | Mostafa Dehghani, Samira Abnar, Jaap Kamps |