Shamsuddeen Muhammad

I'm a Research Scientist in Natural Language Processing at Bayero University, Kano, currently working on low-resource African languages NLP. I previously did research at LIAAD-INESC TEC Porto working with Pavel Brazdil and Alípio Jorge. My work focuses on developing NLP solutions and resources for African languages, particularly in areas like machine translation, sentiment analysis, and hate speech detection. I'm passionate about democratizing NLP technology for African languages through the creation of high-quality datasets and development of effective models. I've helped create several important benchmarks and resources including AfriSenti, MasakhaNER, NaijaSenti, and HaVQA. My research aims to address the challenges of limited data availability and computational resources for African languages.

I actively collaborate with researchers across Africa and globally through initiatives like Masakhane to build capacity and advance African NLP. My work spans multiple areas including cross-lingual transfer learning, multilingual models, and evaluation metrics for low-resource scenarios. I'm particularly interested in making NLP technologies more accessible and effective for the hundreds of millions of African language speakers.

Publications

AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages

AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages

Shamsuddeen Hassan Muhammad, Idris Abdulmumin, A. Ayele, David Ifeoluwa Adelani, I. Ahmad, Saminu Mohammad Aliyu, Nelson Odhiambo Onyango, Lilian D. A. Wanzare, Samuel Rutunda, L. J. Aliyu, E. Alemneh, Oumaima Hourrane, Hagos Tesfahun Gebremichael, Elyas Abdi Ismail, Meriem Beloucif, Ebrahim Chekol Jibril, Andiswa Bukula, Rooweither Mabuya, Salomey Osei, Abigail Oppong, Tadesse Destaw Belay, Tadesse Kebede Guge, T. Asfaw, C. Chukwuneke, Paul Rottger, Seid Muhie Yimam, N. Ousidhoum

AFRIDOC-MT: Document-level MT Corpus for African Languages

AFRIDOC-MT: Document-level MT Corpus for African Languages

Jesujoba Oluwadara Alabi, Israel Abebe Azime, Miaoran Zhang, Cristina España-Bonet, Rachel Bawden, Dawei Zhu, David Ifeoluwa Adelani, Clement Odoje, Idris Akinade, Iffat Maab, Davis David, Shamsuddeen Hassan Muhammad, Neo Putini, David O. Ademuyiwa, Andrew Caines, Dietrich Klakow

Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages

Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages

Edward Bayes, Israel Abebe Azime, Jesujoba Oluwadara Alabi, Jonas Kgomo, Tyna Eloundou, Elizabeth Proehl, Kai Chen, Imaan Khadir, Naome A. Etori, Shamsuddeen Hassan Muhammad, Choice Mpanza, Igneciah Pocia Thete, D. Klakow, David Ifeoluwa Adelani

arXiv.org 2024

Correcting FLORES Evaluation Dataset for Four African Languages

Correcting FLORES Evaluation Dataset for Four African Languages

Idris Abdulmumin, Sthembiso Mkhwanazi, Mahlatse S Mbooi, Shamsuddeen Hassan Muhammad, I. Ahmad, Neo Putini, Miehleketo Mathebula, M. Shingange, T. Gwadabe, Vukosi Marivate

Conference on Machine Translation 2024

Mitigating Translationese in Low-resource Languages: The Storyboard Approach

Mitigating Translationese in Low-resource Languages: The Storyboard Approach

Garry Kuwanto, E. Urua, Priscilla Amuok, Shamsuddeen Hassan Muhammad, Anuoluwapo Aremu, V. Otiende, Loice Emma Nanyanga, T. Nyoike, A. D. Akpan, Nsima Ab Udouboh, Idongesit Udeme Archibong, Idara Effiong Moses, Ifeoluwatayo A. Ige, Benjamin Ayoade Ajibade, Olumide Benjamin Awokoya, Idris Abdulmumin, Saminu Mohammad Aliyu, R. Iro, I. Ahmad, Deontae Smith, Praise-EL Michaels, David Ifeoluwa Adelani, Derry Tanti Wijaya, Anietie U Andy

International Conference on Language Resources and Evaluation 2024

BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages

BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages

Junho Myung, Nayeon Lee, Yi Zhou, Jiho Jin, Rifki Afina Putri, Dimosthenis Antypas, Hsuvas Borkakoty, Eunsu Kim, Carla Pérez-Almendros, A. Ayele, V'ictor Guti'errez-Basulto, Yazm'in Ib'anez-Garc'ia, Hwaran Lee, Shamsuddeen Hassan Muhammad, Kiwoong Park, A. Rzayev, Nina White, Seid Muhie Yimam, Mohammad Taher Pilehvar, N. Ousidhoum, José Camacho-Collados, Alice Oh

arXiv.org 2024

IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models

IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models

David Ifeoluwa Adelani, Jessica Ojo, Israel Abebe Azime, Zhuang Yun Jian, Jesujoba Oluwadara Alabi, Xuanli He, Millicent Ochieng, Sara Hooker, Andiswa Bukula, En-Shiun Annie Lee, Chiamaka Chukwuneke, Happy Buzaaba, Blessing K. Sibanda, Godson Kalipe, Jonathan Mukiibi, Salomon Kabongo KABENAMUALU, Foutse Yuehgoh, M. Setaka, Lolwethu Ndolela, N. Odu, Rooweither Mabuya, Shamsuddeen Hassan Muhammad, Salomey Osei, Sokhar Samb, Tadesse Kebede Guge, Pontus Stenetorp

arXiv.org 2024

SemEval Task 1: Semantic Textual Relatedness for African and Asian Languages

N. Ousidhoum, Shamsuddeen Hassan Muhammad, Mohamed Abdalla, Idris Abdulmumin, I. Ahmad, Sanchit Ahuja, Alham Fikri Aji, Vladimir Araujo, Meriem Beloucif, Christine de Kock, Oumaima Hourrane, Manish Shrivastava, T. Solorio, Nirmal Surange, Krishnapriya Vishnubhotla, Seid Muhie Yimam, Saif Mohammad

International Workshop on Semantic Evaluation 2024

SemRel2024: A Collection of Semantic Textual Relatedness Datasets for 14 Languages

SemRel2024: A Collection of Semantic Textual Relatedness Datasets for 14 Languages

N. Ousidhoum, Shamsuddeen Hassan Muhammad, Mohamed Abdalla, Idris Abdulmumin, I. Ahmad, Sanchit Ahuja, Alham Fikri Aji, Vladimir Araujo, A. Ayele, Pavan Baswani, Meriem Beloucif, Christian Biemann, Sofia Bourhim, Christine de Kock, Genet Shanko Dekebo, Oumaima Hourrane, Gopichand Kanumolu, Lokesh Madasu, Samuel Rutunda, Manish Shrivastava, T. Solorio, Nirmal Surange, Hailegnaw Getaneh Tilaye, Krishnapriya Vishnubhotla, Genta Indra Winata, Seid Muhie Yimam, Saif Mohammad

Annual Meeting of the Association for Computational Linguistics 2024

Analyzing COVID-19 Vaccination Sentiments in Nigerian Cyberspace: Insights from a Manually Annotated Twitter Dataset

I. Ahmad, L. J. Aliyu, Abubakar Auwal Khalid, S. M. Aliyu, Shamsuddeen Hassan Muhammad, Idris Abdulmumin, B.M. Abduljalil, Bello Shehu Bello, Amina Imam Abubakar

arXiv.org 2024

Leveraging Closed-Access Multilingual Embedding for Automatic Sentence Alignment in Low Resource Languages

Leveraging Closed-Access Multilingual Embedding for Automatic Sentence Alignment in Low Resource Languages

Idris Abdulmumin, Auwal Abubakar Khalid, Shamsuddeen Hassan Muhammad, I. Ahmad, L. J. Aliyu, Babangida Sani, B.M. Abduljalil, Sani Ahmad Hassan

arXiv.org 2023

AfriMTE and AfriCOMET: Enhancing COMET to Embrace Under-resourced African Languages

AfriMTE and AfriCOMET: Enhancing COMET to Embrace Under-resourced African Languages

Jiayi Wang, David Ifeoluwa Adelani, Sweta Agrawal, Marek Masiak, Ricardo Rei, Eleftheria Briakou, Marine Carpuat, Xuanli He, Sofia Bourhim, Andiswa Bukula, Muhidin A. Mohamed, Temitayo Olatoye, Tosin P. Adewumi, Hamam Mokayede, Christine Mwase, Wangui Kimotho, Foutse Yuehgoh, Anuoluwapo Aremu, Jessica Ojo, Shamsuddeen Hassan Muhammad, Salomey Osei, Abdul-Hakeem Omotayo, Chiamaka Chukwuneke, Perez Ogayo, Oumaima Hourrane, Salma El Anigri, Lolwethu Ndolela, Thabiso Mangwana, Shafie Abdi Mohamed, Ayinde Hassan, Oluwabusayo Olufunke Awoyomi, Lama Alkhaled, S. Al-Azzawi, Naome A. Etori, Millicent Ochieng, Clemencia Siro, Samuel Njoroge, Eric Muchiri, Wangari Kimotho, Lyse Naomi Wamba Momo, D. Abolade, Simbiat Ajao, Iyanuoluwa Shode, Ricky Macharm, R. Iro, S. S. Abdullahi, Stephen E. Moore, Bernard Opoku, Zainab Akinjobi, Abeeb Afolabi, Nnaemeka Obiefuna, Onyekachi Raphael Ogbu, Sam Brian, V. Otiende, C. Mbonu, Sakayo Toadoum Sari, Yao Lu, Pontus Stenetorp

North American Chapter of the Association for Computational Linguistics 2023

AfriWOZ: Corpus for Exploiting Cross-Lingual Transfer for Dialogue Generation in Low-Resource, African Languages

AfriWOZ: Corpus for Exploiting Cross-Lingual Transfer for Dialogue Generation in Low-Resource, African Languages

Tosin P. Adewumi, Mofetoluwa Adeyemi, Aremu Anuoluwapo, Bukola Peters, Happy Buzaaba, Oyerinde Samuel, Amina Mardiyyah Rufai, Benjamin Ayoade Ajibade, Tajudeen Gwadabe, Mory Moussou Koulibaly Traore, T. Ajayi, Shamsuddeen Hassan Muhammad, Ahmed Baruwa, Paul Owoicho, Tolúlopé Ògúnrèmí, Phylis Ngigi, Orevaoghene Ahia, Ruqayya Nasir, F. Liwicki, M. Liwicki

IEEE International Joint Conference on Neural Network 2023

HaVQA: A Dataset for Visual Question Answering and Multimodal Research in Hausa Language

HaVQA: A Dataset for Visual Question Answering and Multimodal Research in Hausa Language

Shantipriya Parida, Idris Abdulmumin, Shamsuddeen Hassan Muhammad, Aneesh Bose, Guneet Singh Kohli, I. Ahmad, Ketan Kotwal, Sayan Deb Sarkar, Ondrej Bojar, H. Kakudi

Annual Meeting of the Association for Computational Linguistics 2023

MasakhaPOS: Part-of-Speech Tagging for Typologically Diverse African languages

MasakhaPOS: Part-of-Speech Tagging for Typologically Diverse African languages

Cheikh M. Bamba Dione, David Ifeoluwa Adelani, Peter Nabende, Jesujoba Oluwadara Alabi, Thapelo Sindane, Happy Buzaaba, Shamsuddeen Hassan Muhammad, Chris C. Emezue, Perez Ogayo, Anuoluwapo Aremu, Catherine Gitau, Derguene Mbaye, Jonathan Mukiibi, Blessing K. Sibanda, Bonaventure F. P. Dossou, Andiswa Bukula, Rooweither Mabuya, A. Tapo, Edwin Munkoh-Buabeng, V. M. Koagne, F. Kabore, Amelia Taylor, Godson Kalipe, Tebogo Macucwa, Vukosi Marivate, T. Gwadabe, Mboning Tchiaze Elvis, I. Onyenwe, G. Atindogbé, T. Adelani, Idris Akinade, Olanrewaju Samuel, M. Nahimana, Th'eogene Musabeyezu, Emile Niyomutabazi, Ester Chimhenga, Kudzai Gotosa, Patrick Mizha, Apelete Agbolo, Seydou T. Traoré, C. Uchechukwu, Aliyu Yusuf, M. Abdullahi, D. Klakow

Annual Meeting of the Association for Computational Linguistics 2023

AfriQA: Cross-lingual Open-Retrieval Question Answering for African Languages

AfriQA: Cross-lingual Open-Retrieval Question Answering for African Languages

Odunayo Ogundepo, T. Gwadabe, Clara Rivera, J. Clark, Sebastian Ruder, David Ifeoluwa Adelani, Bonaventure F. P. Dossou, Abdoulahat Diop, Claytone Sikasote, Gilles Hacheme, Happy Buzaaba, Ignatius M Ezeani, Rooweither Mabuya, Salomey Osei, Chris C. Emezue, A. Kahira, Shamsuddeen Hassan Muhammad, Akintunde Oladipo, A. Owodunni, A. Tonja, Iyanuoluwa Shode, Akari Asai, T. Ajayi, Clemencia Siro, Steven Arthur, Mofetoluwa Adeyemi, Orevaoghene Ahia, Aremu Anuoluwapo, O. Awosan, C. Chukwuneke, Bernard Opoku, A. Ayodele, V. Otiende, Christine Mwase, B. Sinkala, Andre Niyongabo Rubungo, Daniel Ajisafe, Emeka Onwuegbuzia, Habib Mbow, Emile Niyomutabazi, Eunice Mukonde, F. I. Lawan, I. Ahmad, Jesujoba Oluwadara Alabi, Martin Namukombo, Mbonu Chinedu, Mofya Phiri, Neo Putini, Ndumiso Mngoma, Priscilla Amuok, R. Iro, Sonia Adhiambo34

Conference on Empirical Methods in Natural Language Processing 2023

HausaNLP at SemEval-2023 Task 10: Transfer Learning, Synthetic Data and Side-information for Multi-level Sexism Classification

HausaNLP at SemEval-2023 Task 10: Transfer Learning, Synthetic Data and Side-information for Multi-level Sexism Classification

Saminu Mohammad Aliyu, Idris Abdulmumin, Shamsuddeen Hassan Muhammad, I. Ahmad, Saheed Abdullahi Salahudeen, Aliyu Yusuf, F. I. Lawan

International Workshop on Semantic Evaluation 2023

MasakhaNEWS: News Topic Classification for African languages

MasakhaNEWS: News Topic Classification for African languages

David Ifeoluwa Adelani, Marek Masiak, Israel Abebe Azime, Jesujoba Oluwadara Alabi, A. Tonja, Christine Mwase, Odunayo Ogundepo, Bonaventure F. P. Dossou, Akintunde Oladipo, Doreen Nixdorf, Chris C. Emezue, S. Al-Azzawi, Blessing K. Sibanda, Davis David, Lolwethu Ndolela, Jonathan Mukiibi, T. Ajayi, Tatiana Moteu Ngoli, B. Odhiambo, A. Owodunni, Nnaemeka Obiefuna, Shamsuddeen Hassan Muhammad, S. S. Abdullahi, M. Yigezu, T. Gwadabe, Idris Abdulmumin, Mahlet Taye Bame, Oluwabusayo Olufunke Awoyomi, Iyanuoluwa Shode, T. Adelani, Habiba Abdulganiy Kailani, Abdul-Hakeem Omotayo, Adetola Adeeko, Afolabi Abeeb, Anuoluwapo Aremu, Olanrewaju Samuel, Clemencia Siro, Wangari Kimotho, Onyekachi Raphael Ogbu, C. Mbonu, C. Chukwuneke, Samuel Fanijo, Jessica Ojo, Oyinkansola F. Awosan, Tadesse Kebede Guge, Sakayo Toadoum Sari, Pamela Nyatsine, Freedmore Sidume, Oreen Yousuf, Mardiyyah Oduwole, Ussen Kimanuka, Kanda Patrick Tshinu, Thina Diko, Siyanda Nxakama, Abdulmejid Tuni Johar, Sinodos Gebre, Muhidin A. Mohamed, Shafie Abdi Mohamed, Fuad Mire Hassan, Moges Ahmed Mehamed, Evrard Ngabire, Pontus Stenetorp

International Joint Conference on Natural Language Processing 2023

SemEval-2023 Task 12: Sentiment Analysis for African Languages (AfriSenti-SemEval)

SemEval-2023 Task 12: Sentiment Analysis for African Languages (AfriSenti-SemEval)

Shamsuddeen Hassan Muhammad, Idris Abdulmumin, Seid Muhie Yimam, David Ifeoluwa Adelani, I. Ahmad, N. Ousidhoum, A. Ayele, Saif M. Mohammad, Meriem Beloucif

International Workshop on Semantic Evaluation 2023

The African Stopwords project: curating stopwords for African languages

The African Stopwords project: curating stopwords for African languages

Chris C. Emezue, H. Nigatu, Cynthia Thinwa, He Zhou, Shamsuddeen Hassan Muhammad, Lerato Louis, Idris Abdulmumin, S. Oyerinde, Benjamin Ayoade Ajibade, Olanrewaju Samuel, Oviawe Joshua, Emeka Onwuegbuzia, Handel Emezue, Ifeoluwatayo A. Ige, A. Tonja, C. Chukwuneke, Bonaventure F. P. Dossou, Naome A. Etori, Mbonu Chinedu Emmanuel, Oreen Yousuf, Kaosarat Aina, Davis David

arXiv.org 2023

AfriSenti: A Twitter Sentiment Analysis Benchmark for African Languages

AfriSenti: A Twitter Sentiment Analysis Benchmark for African Languages

Shamsuddeen Hassan Muhammad, Idris Abdulmumin, A. Ayele, N. Ousidhoum, David Ifeoluwa Adelani, Seid Muhie Yimam, I. Ahmad, Meriem Beloucif, Saif M. Mohammad, Sebastian Ruder, Oumaima Hourrane, P. Brazdil, Felermino D'ario M'ario Ant'onio Ali, Davis C. Davis, Salomey Osei, Bello Shehu Bello, Falalu Ibrahim, T. Gwadabe, Samuel Rutunda, Tadesse Destaw Belay, Wendimu Baye Messelle, Hailu Beshada Balcha, S. Chala, Hagos Tesfahun Gebremichael, Bernard Opoku, Steven Arthur

Conference on Empirical Methods in Natural Language Processing 2023

HERDPhobia: A Dataset for Hate Speech against Fulani in Nigeria

HERDPhobia: A Dataset for Hate Speech against Fulani in Nigeria

Saminu Mohammad Aliyu, G. Wajiga, M. Murtala, Shamsuddeen Hassan Muhammad, Idris Abdulmumin, I. Ahmad

arXiv.org 2022

BLOOM: A 176B-Parameter Open-Access Multilingual Language Model

BLOOM: A 176B-Parameter Open-Access Multilingual Language Model

Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ili'c, Daniel Hesslow, Roman Castagn'e, A. Luccioni, François Yvon, Matthias Gallé, J. Tow, Alexander M. Rush, Stella Biderman, Albert Webson, Pawan Sasanka Ammanamanchi, Thomas Wang, Benoît Sagot, Niklas Muennighoff, Albert Villanova del Moral, Olatunji Ruwase, Rachel Bawden, Stas Bekman, Angelina McMillan-Major, Iz Beltagy, Huu Nguyen, Lucile Saulnier, Samson Tan, Pedro Ortiz Suarez, Victor Sanh, Hugo Laurenccon, Yacine Jernite, Julien Launay, Margaret Mitchell, Colin Raffel, Aaron Gokaslan, Adi Simhi, Aitor Soroa Etxabe, Alham Fikri Aji, Amit Alfassy, Anna Rogers, Ariel Kreisberg Nitzav, Canwen Xu, Chenghao Mou, Chris C. Emezue, Christopher Klamm, Colin Leong, Daniel Alexander van Strien, David Ifeoluwa Adelani, Dragomir R. Radev, E. G. Ponferrada, Efrat Levkovizh, Ethan Kim, Eyal Natan, F. Toni, Gérard Dupont, Germán Kruszewski, Giada Pistilli, Hady ElSahar, Hamza Benyamina, H. Tran, Ian Yu, Idris Abdulmumin, Isaac Johnson, Itziar Gonzalez-Dios, Javier de la Rosa, Jenny Chim, Jesse Dodge, Jian Zhu, Jonathan Chang, Jorg Frohberg, Josephine Tobing, J. Bhattacharjee, Khalid Almubarak, Kimbo Chen, Kyle Lo, Leandro von Werra, Leon Weber, Long Phan, Loubna Ben Allal, Ludovic Tanguy, Manan Dey, M. Muñoz, Maraim Masoud, María Grandury, Mario vSavsko, Max Huang, Maximin Coavoux, Mayank Singh, Mike Tian-Jian Jiang, Minh Chien Vu, M. A. Jauhar, Mustafa Ghaleb, Nishant Subramani, Nora Kassner, Nurulaqilla Khamis, Olivier Nguyen, Omar Espejel, Ona de Gibert, Paulo Villegas, Peter Henderson, Pierre Colombo, Priscilla Amuok, Quentin Lhoest, Rheza Harliman, Rishi Bommasani, R. L'opez, Rui Ribeiro, Salomey Osei, S. Pyysalo, Sebastian Nagel, Shamik Bose, Shamsuddeen Hassan Muhammad, Shanya Sharma, S. Longpre, Somaieh Nikpoor, S. Silberberg, S. Pai, S. Zink, Tiago Timponi Torrent, Timo Schick, Tristan Thrush, V. Danchev, Vassilina Nikoulina, Veronika Laippala, Violette Lepercq, V. Prabhu, Zaid Alyafeai, Zeerak Talat, Arun Raja, Benjamin Heinzerling, Chenglei Si, Elizabeth Salesky, Sabrina J. Mielke, Wilson Y. Lee, Abheesht Sharma, Andrea Santilli, Antoine Chaffin, Arnaud Stiegler, Debajyoti Datta, Eliza Szczechla, Gunjan Chhablani, Han Wang, Harshit Pandey, Hendrik Strobelt, Jason Alan Fries, Jos Rozen, Leo Gao, Lintang Sutawika, M Saiful Bari, Maged S. Al-Shaibani, Matteo Manica, Nihal V. Nayak, Ryan Teehan, Samuel Albanie, Sheng Shen, Srulik Ben-David, Stephen H. Bach, Taewoon Kim, T. Bers, Thibault Févry, Trishala Neeraj, Urmish Thakker, Vikas Raunak, Xiang Tang, Zheng-Xin Yong, Zhiqing Sun, Shaked Brody, Y. Uri, Hadar Tojarieh, Adam Roberts, Hyung Won Chung, Jaesung Tae, Jason Phang, Ofir Press, Conglong Li, D. Narayanan, Hatim Bourfoune, J. Casper, Jeff Rasley, Max Ryabinin, Mayank Mishra, Minjia Zhang, Mohammad Shoeybi, Myriam Peyrounette, N. Patry, Nouamane Tazi, Omar Sanseviero, Patrick von Platen, Pierre Cornette, Pierre Franccois Lavall'ee, R. Lacroix, Samyam Rajbhandari, Sanchit Gandhi, Shaden Smith, S. Requena, Suraj Patil, Tim Dettmers, Ahmed Baruwa, Amanpreet Singh, Anastasia Cheveleva, Anne-Laure Ligozat, Arjun Subramonian, Aur'elie N'ev'eol, Charles Lovering, Daniel H Garrette, D. Tunuguntla, Ehud Reiter, Ekaterina Taktasheva, E. Voloshina, Eli Bogdanov, Genta Indra Winata, Hailey Schoelkopf, Jan-Christoph Kalo, Jekaterina Novikova, J. Forde, Xiangru Tang, Jungo Kasai, Ken Kawamura, Liam Hazan, Marine Carpuat, Miruna Clinciu, Najoung Kim, Newton Cheng, O. Serikov, Omer Antverg, Oskar van der Wal, Rui Zhang, Ruochen Zhang, Sebastian Gehrmann, Shachar Mirkin, S. Pais, Tatiana Shavrina, Thomas Scialom, Tian Yun, Tomasz Limisiewicz, Verena Rieser, Vitaly Protasov, V. Mikhailov, Yada Pruksachatkun, Yonatan Belinkov, Zachary Bamberger, Zdenvek Kasner, Zdeněk Kasner, A. Pestana, A. Feizpour, Ammar Khan, Amy Faranak, A. Santos, Anthony Hevia, Antigona Unldreaj, Arash Aghagol, Arezoo Abdollahi, A. Tammour, A. HajiHosseini, Bahareh Behroozi, Benjamin Ayoade Ajibade, B. Saxena, Carlos Muñoz Ferrandis, Danish Contractor, D. Lansky, Davis David, Douwe Kiela, D. A. Nguyen, Edward Tan, Emi Baylor, Ezinwanne Ozoani, F. Mirza, Frankline Ononiwu, Habib Rezanejad, H.A. Jones, Indrani Bhattacharya, Irene Solaiman, Irina Sedenko, Isar Nejadgholi, J. Passmore, Joshua Seltzer, Julio Bonis Sanz, Karen Fort, Lívia Dutra, Mairon Samagaio, Maraim Elbadri, Margot Mieskes, Marissa Gerchick, Martha Akinlolu, Michael McKenna, Mike Qiu, M. Ghauri, Mykola Burynok, Nafis Abrar, Nazneen Rajani, Nour Elkott, N. Fahmy, Olanrewaju Samuel, Ran An, R. Kromann, Ryan Hao, S. Alizadeh, Sarmad Shubber, Silas L. Wang, Sourav Roy, S. Viguier, Thanh-Cong Le, Tobi Oyebade, T. Le, Yoyo Yang, Zach Nguyen, Abhinav Ramesh Kashyap, Alfredo Palasciano, A. Callahan, Anima Shukla, Antonio Miranda-Escalada, A. Singh, Benjamin Beilharz, Bo Wang, C. Brito, Chenxi Zhou, Chirag Jain, Chuxin Xu, Clémentine Fourrier, Daniel Le'on Perin'an, Daniel Molano, Dian Yu, Enrique Manjavacas, Fabio Barth, Florian Fuhrimann, Gabriel Altay, Giyaseddin Bayrak, Gully Burns, Helena U. Vrabec, I. Bello, Isha Dash, J. Kang, John Giorgi, Jonas Golde, J. Posada, Karthi Sivaraman, Lokesh Bulchandani, Lu Liu, Luisa Shinzato, Madeleine Hahn de Bykhovetz, Maiko Takeuchi, Marc Pàmies, M. A. Castillo, Marianna Nezhurina, Mario Sanger, M. Samwald, Michael Cullan, Michael Weinberg, M. Wolf, Mina Mihaljcic, Minna Liu, M. Freidank, Myungsun Kang, Natasha Seelam, N. Dahlberg, N. Broad, N. Muellner, Pascale Fung, Patricia Haller, Patrick Haller, R. Eisenberg, Robert Martin, Rodrigo Canalli, Rosaline Su, Ruisi Su, Samuel Cahyawijaya, Samuele Garda, Shlok S Deshmukh, Shubhanshu Mishra, Sid Kiblawi, Simon Ott, Sinee Sang-aroonsiri, Srishti Kumar, Stefan Schweter, S. Bharati, Tanmay Laud, Théo Gigant, Tomoya Kainuma, Wojciech Kusa, Yanis Labrak, Yashasvi Bajaj, Y. Venkatraman, Yifan Xu, Ying Xu, Yu Xu, Z. Tan, Zhongli Xie, Zifan Ye, M. Bras, Younes Belkada, Thomas Wolf

arXiv.org 2022

MasakhaNER 2.0: Africa-centric Transfer Learning for Named Entity Recognition

MasakhaNER 2.0: Africa-centric Transfer Learning for Named Entity Recognition

David Ifeoluwa Adelani, Graham Neubig, Sebastian Ruder, Shruti Rijhwani, Michael Beukman, Chester Palen-Michel, Constantine Lignos, Jesujoba Oluwadara Alabi, Shamsuddeen Hassan Muhammad, Peter Nabende, Cheikh M. Bamba Dione, Andiswa Bukula, Rooweither Mabuya, Bonaventure F. P. Dossou, Blessing K. Sibanda, Happy Buzaaba, Jonathan Mukiibi, Godson Kalipe, Derguene Mbaye, Amelia Taylor, F. Kabore, Chris C. Emezue, Anuoluwapo Aremu, Perez Ogayo, C. Gitau, Edwin Munkoh-Buabeng, V. M. Koagne, A. Tapo, Tebogo Macucwa, Vukosi Marivate, Elvis Mboning, T. Gwadabe, Tosin P. Adewumi, Orevaoghene Ahia, J. Nakatumba‐Nabende, Neo L. Mokono, Ignatius M Ezeani, C. Chukwuneke, Mofetoluwa Adeyemi, Gilles Hacheme, Idris Abdulmumin, Odunayo Ogundepo, Oreen Yousuf, Tatiana Moteu Ngoli, D. Klakow

Conference on Empirical Methods in Natural Language Processing 2022

Separating Grains from the Chaff: Using Data Filtering to Improve Multilingual Translation for Low-Resourced African Languages

Separating Grains from the Chaff: Using Data Filtering to Improve Multilingual Translation for Low-Resourced African Languages

Idris Abdulmumin, Michael Beukman, Jesujoba Oluwadara Alabi, Chris C. Emezue, Everlyn Asiko, Tosin P. Adewumi, Shamsuddeen Hassan Muhammad, Mofetoluwa Adeyemi, Oreen Yousuf, Sahib Singh, T. Gwadabe

Conference on Machine Translation 2022

Semi-Automatic Approaches for Exploiting Shifter Patterns in Domain-Specific Sentiment Analysis

P. Brazdil, Shamsuddeen Hassan Muhammad, F. Oliveira, João Paulo Cordeiro, Fátima Silva, Purificação Silvano, Antonio Leal

Mathematics 2022

BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus

BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus

Josh Meyer, David Ifeoluwa Adelani, Edresson Casanova, A. Oktem, Daniel Whitenack Julian Weber, Salomon Kabongo KABENAMUALU, Elizabeth Salesky, Iroro Orife, Colin Leong, Perez Ogayo, Chris C. Emezue, Jonathan Mukiibi, Salomey Osei, Apelete Agbolo, Victor Akinode, Bernard Opoku, S. Olanrewaju, Jesujoba Oluwadara Alabi, Shamsuddeen Hassan Muhammad

Interspeech 2022

A Few Thousand Translations Go a Long Way! Leveraging Pre-trained Models for African News Translation

A Few Thousand Translations Go a Long Way! Leveraging Pre-trained Models for African News Translation

David Ifeoluwa Adelani, Jesujoba Oluwadara Alabi, Angela Fan, Julia Kreutzer, Xiaoyu Shen, Machel Reid, Dana Ruiter, D. Klakow, Peter Nabende, Ernie Chang, T. Gwadabe, Freshia Sackey, Bonaventure F. P. Dossou, Chris C. Emezue, Colin Leong, Michael Beukman, Shamsuddeen Hassan Muhammad, Guyo Dub Jarso, Oreen Yousuf, Andre Niyongabo Rubungo, Gilles Hacheme, Eric Peter Wairagala, Muhammad Umair Nasir, Benjamin Ayoade Ajibade, T. Ajayi, Yvonne Wambui Gitau, Jade Z. Abbott, Mohamed Ahmed, Millicent Ochieng, Anuoluwapo Aremu, Perez Ogayo, Jonathan Mukiibi, F. Kabore, Godson Kalipe, Derguene Mbaye, A. Tapo, V. M. Koagne, Edwin Munkoh-Buabeng, Valencia Wagner, Idris Abdulmumin, Ayodele Awokoya, Happy Buzaaba, Blessing K. Sibanda, Andiswa Bukula, Sam Manthalu

North American Chapter of the Association for Computational Linguistics 2022

Hausa Visual Genome: A Dataset for Multi-Modal English to Hausa Machine Translation

Hausa Visual Genome: A Dataset for Multi-Modal English to Hausa Machine Translation

Idris Abdulmumin, S. Dash, Musa Abdullahi Dawud, Shantipriya Parida, Shamsuddeen Hassan Muhammad, I. Ahmad, Subhadarshi Panda, Ondrej Bojar, B. Galadanci, Bello Shehu Bello

International Conference on Language Resources and Evaluation 2022

AfriWOZ: Corpus for Exploiting Cross-Lingual Transferability for Generation of Dialogues in Low-Resource, African Languages

AfriWOZ: Corpus for Exploiting Cross-Lingual Transferability for Generation of Dialogues in Low-Resource, African Languages

Tosin P. Adewumi, Mofetoluwa Adeyemi, Aremu Anuoluwapo, Bukola Peters, Happy Buzaaba, Oyerinde Samuel, Amina Mardiyyah Rufai, Benjamin Ayoade Ajibade, Tajudeen Gwadabe, M. Traore, T. Ajayi, Shamsuddeen Hassan Muhammad, Ahmed Baruwa, Paul Owoicho, Tolúlopé Ògúnrèmí, Phylis Ngigi, Orevaoghene Ahia, Ruqayya Nasir, F. Liwicki, M. Liwicki

Quantity vs. Quality of Monolingual Source Data in Automatic Text Translation: Can It Be Too Little If It Is Too Good?

Quantity vs. Quality of Monolingual Source Data in Automatic Text Translation: Can It Be Too Little If It Is Too Good?

Idris Abdulmumin, B. Galadanci, Shamsuddeen Hassan Muhammad, Garba Aliyu

2022 IEEE Nigeria 4th International Conference on Disruptive Technologies for Sustainable Development (NIGERCON) 2022

NaijaSenti: A Nigerian Twitter Sentiment Corpus for Multilingual Sentiment Analysis

NaijaSenti: A Nigerian Twitter Sentiment Corpus for Multilingual Sentiment Analysis

Shamsuddeen Hassan Muhammad, David Ifeoluwa Adelani, I. Ahmad, Idris Abdulmumin, Bello Shehu Bello, M. Choudhury, Chris C. Emezue, Anuoluwapo Aremu, Saheed Abdul, P. Brazdil

International Conference on Language Resources and Evaluation 2022

Deep Sequence Models for Text Classification Tasks

Deep Sequence Models for Text Classification Tasks

S. S. Abdullahi, Su Yiming, Shamsuddeen Hassan Muhammad, A. Mustapha, Ahmad Muhammad Aminu, Abdulkadir Abdullahi, Musa Bello, Saminu Mohammad Aliyu

2021 International Conference on Electrical, Communication, and Computer Engineering (ICECCE) 2021

Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets

Isaac Caswell, Julia Kreutzer, Lisa Wang, Ahsan Wahab, D. Esch, Nasanbayar Ulzii-Orshikh, A. Tapo, Nishant Subramani, Artem Sokolov, Claytone Sikasote, Monang Setyawan, Supheakmungkol Sarin, Sokhar Samb, B. Sagot, Clara Rivera, Annette Rios Gonzales, Isabel Papadimitriou, Salomey Osei, Pedro Ortiz Suarez, Iroro Orife, Kelechi Ogueji, Andre Niyongabo Rubungo, Toan Q. Nguyen, Mathias Muller, A. Muller, Shamsuddeen Hassan Muhammad, N. Muhammad, Ayanda Mnyakeni, Jamshidbek Mirzakhalov, Tapiwanashe Matangira, Colin Leong, Nze Lawson, Sneha Kudugunta, Yacine Jernite, M. Jenny, Orhan Firat, Bonaventure F. P. Dossou, Sakhile Dlamini, Nisansa de Silva, Sakine cCabuk Balli, Stella Biderman, A. Battisti, Ahmed Baruwa, Ankur Bapna, P. Baljekar, Israel Abebe Azime, Ayodele Awokoya, Duygu Ataman, Orevaoghene Ahia, Oghenefego Ahia, Sweta Agrawal, Mofetoluwa Adeyemi

Transactions of the Association for Computational Linguistics 2021

MasakhaNER: Named Entity Recognition for African Languages

David Ifeoluwa Adelani, Jade Z. Abbott, Graham Neubig, Daniel D'souza, Julia Kreutzer, Constantine Lignos, Chester Palen-Michel, Happy Buzaaba, Shruti Rijhwani, Sebastian Ruder, Stephen Mayhew, Israel Abebe Azime, Shamsuddeen Hassan Muhammad, Chris C. Emezue, J. Nakatumba‐Nabende, Perez Ogayo, Anuoluwapo Aremu, Catherine Gitau, Derguene Mbaye, Jesujoba Oluwadara Alabi, Seid Muhie Yimam, T. Gwadabe, I. Ezeani, Andre Niyongabo Rubungo, Jonathan Mukiibi, V. Otiende, Iroro Orife, Davis David, Samba Ngom, Tosin P. Adewumi, Paul Rayson, Mofetoluwa Adeyemi, Gerald Muriuki, E. Anebi, C. Chukwuneke, N. Odu, Eric Peter Wairagala, S. Oyerinde, Clemencia Siro, Tobius Saul Bateesa, Temilola Oloyede, Yvonne Wambui, Victor Akinode, Deborah Nabagereka, Maurice Katusiime, Ayodele Awokoya, Mouhamadane Mboup, Dibora Gebreyohannes, Henok Tilaye, Kelechi Nwaike, Degaga Wolde, A. Faye, Blessing K. Sibanda, Orevaoghene Ahia, Bonaventure F. P. Dossou, Kelechi Ogueji, T. Diop, A. Diallo, Adewale Akinfaderin, T. Marengereke, Salomey Osei

Transactions of the Association for Computational Linguistics 2021

Participatory Research for Low-resourced Machine Translation: A Case Study in African Languages

Participatory Research for Low-resourced Machine Translation: A Case Study in African Languages

W. Nekoto, Vukosi Marivate, T. Matsila, Timi E. Fasubaa, T. Kolawole, T. Fagbohungbe, S. Akinola, Shamsuddeen Hassan Muhammad, Salomon Kabongo KABENAMUALU, Salomey Osei, Sackey Freshia, Andre Niyongabo Rubungo, Ricky Macharm, Perez Ogayo, Orevaoghene Ahia, Musie Meressa, Mofetoluwa Adeyemi, Masabata Mokgesi-Selinga, Lawrence Okegbemi, L. Martinus, Kolawole Tajudeen, Kevin Degila, Kelechi Ogueji, Kathleen Siminyu, Julia Kreutzer, Jason Webster, Jamiil Toure Ali, Jade Z. Abbott, Iroro Orife, I. Ezeani, Idris Abdulkabir Dangana, H. Kamper, Hady ElSahar, Goodness Duru, Ghollah Kioko, Espoir Murhabazi, Elan Van Biljon, Daniel Whitenack, Christopher Onyefuluchi, Chris C. Emezue, Bonaventure F. P. Dossou, Blessing K. Sibanda, B. Bassey, A. Olabiyi, A. Ramkilowan, A. Oktem, Adewale Akinfaderin, Abdallah Bashir

Findings 2020

A Survey on Machine Learning Techniques in Movie Revenue Prediction

I. Ahmad, Adeela Abu Bakar, Mohd Ridzwan Yaakub, Shamsuddeen Hassan Muhammad

SN Computer Science 2020

Incremental Approach for Automatic Generation of Domain-Specific Sentiment Lexicon

Shamsuddeen Hassan Muhammad, P. Brazdil, A. Jorge

European Conference on Information Retrieval 2020

A Form of List Viterbi Algorithm for Decoding Convolutional Codes

A Form of List Viterbi Algorithm for Decoding Convolutional Codes

Shamsuddeen Hassan Muhammad, A. Mustapha

U Porto Journal of Engineering 2018

HausaHate: An Expert Annotated Corpus for Hausa Hate Speech Detection

HausaHate: An Expert Annotated Corpus for Hausa Hate Speech Detection

F. Vargas, Samuel Guimarães, Shamsuddeen Hassan Muhammad, Diego Alves, I. Ahmad, Idris Abdulmumin, Diallo Mohamed, Thiago A. S. Pardo, Fabrício Benevenuto

WOAH 2024

Findings of WMT2024 English-to-Low Resource Multimodal Translation Task

Findings of WMT2024 English-to-Low Resource Multimodal Translation Task

Shantipriya Parida, Ondrej Bojar, Idris Abdulmumin, Shamsuddeen Hassan Muhammad, I. Ahmad

Conference on Machine Translation 2024

SemEval Task 1: Semantic Textual Relatedness for African and Asian Languages

SemEval Task 1: Semantic Textual Relatedness for African and Asian Languages

N. Ousidhoum, Shamsuddeen Hassan Muhammad, Mohamed Abdalla, Idris Abdulmumin, I. Ahmad, Sanchit Ahuja, Alham Fikri Aji, Vladimir Araujo, Meriem Beloucif, Christine de Kock, Oumaima Hourrane, Manish Shrivastava, T. Solorio, Nirmal Surange, Krishnapriya Vishnubhotla, Seid Muhie Yimam, Saif Mohammad

SemEval@NAACL 2024

Combining Symbolic and Deep Learning Approaches for Sentiment Analysis

Shamsuddeen Hassan Muhammad, P. Brazdil, A. Jorge

Compendium of Neurosymbolic Artificial Intelligence 2023

A FRI S ENTI : A B ENCHMARK T WITTER S ENTIMENT A NALYSIS D ATASET FOR A FRICAN L ANGUAGES

A FRI S ENTI : A B ENCHMARK T WITTER S ENTIMENT A NALYSIS D ATASET FOR A FRICAN L ANGUAGES

Shamsuddeen Hassan Muhammad, Idris Abdulmumin, A. Ayele, N. Ousidhoum, David Ifeoluwa Adelani, Seid Muhie Yimam, Meriem Beloucif, Saif M. Mohammad, Sebastian Ruder, Oumaima Hourrane, P. Brazdil, Felermino M. D. A. Ali, Davis David, Salomey Osei, Bello Shehu Bello, Falalu Ibrahim, T. Gwadabe, Samuel Rutunda, Tadesse Destaw Belay, Wendimu Baye Messelle, Hailu Beshada Balcha, S. Chala, Hagos Tesfahun Gebremichael, Bernard Opoku, Steven Arthur

Symbolic Versus Deep Learning Techniques for Explainable Sentiment Analysis

Shamsuddeen Hassan Muhammad, P. Brazdil, A. Jorge

Portuguese Conference on Artificial Intelligence 2023

AfriMTE and AfriCOMET: Empowering COMET to Embrace Under-resourced African Languages

AfriMTE and AfriCOMET: Empowering COMET to Embrace Under-resourced African Languages

Jiayi Wang, David Ifeoluwa Adelani, Sweta Agrawal, Ricardo Rei, Eleftheria Briakou, Marine Carpuat, Marek Masiak, Xuanli He, Sofia Bourhim, Andiswa Bukula, Muhidin A. Mohamed, Temitayo Olatoye, Hamam Mokayede, Christine Mwase, Wangui Kimotho, Foutse Yuehgoh, Anuoluwapo Aremu, Jessica Ojo, Shamsuddeen Hassan Muhammad, Salomey Osei, Abdul-Hakeem Omotayo, Chiamaka Chukwuneke, Perez Ogayo, Oumaima Hourrane, Salma El Anigri, Lolwethu Ndolela, Thabiso Mangwana, Shafie Abdi Mohamed, Ayinde Hassan, Oluwabusayo Olufunke Awoyomi, Lama Alkhaled, S. Al-Azzawi, Naome A. Etori, Millicent Ochieng, Clemencia Siro, Samuel Njoroge, Eric Muchiri, Wangari Kimotho, Lyse Naomi Wamba Momo, D. Abolade, Simbiat Ajao, Tosin P. Adewumi, Iyanuoluwa Shode, Ricky Macharm, R. Iro, S. S. Abdullahi, Stephen E. Moore, Bernard Opoku, Zainab Akinjobi, Abeeb Afolabi, Nnaemeka Obiefuna, Onyekachi Raphael Ogbu, Sam Brian, V. Otiende, C. Mbonu, Sakayo Toadoum Sari, Pontus Stenetorp

arXiv.org 2023

Ìtàkúròso: Exploiting Cross-Lingual Transferability for Natural Language Generation of Dialogues in Low-Resource, African Languages

Ìtàkúròso: Exploiting Cross-Lingual Transferability for Natural Language Generation of Dialogues in Low-Resource, African Languages

Tosin P. Adewumi, Mofetoluwa Adeyemi, Aremu Anuoluwapo, Bukola Peters, Happy Buzaaba, Oyerinde Samuel, Amina Mardiyyah Rufai, Benjamin Ayoade Ajibade, Tajudeen Gwadabe, M. Traore, T. Ajayi, Shamsuddeen Hassan Muhammad, Ahmed Baruwa, Paul Owoicho, Tolúlopé Ògúnrèmí, Phylis Ngigi, Orevaoghene Ahia, Ruqayya Nasir, F. Liwicki, M. Liwicki

arXiv.org 2022

Phone-ing it in: Towards Flexible, Multi-Modal Language Model Training using Phonetic Representations of Data

Phone-ing it in: Towards Flexible, Multi-Modal Language Model Training using Phonetic Representations of Data

Shamsuddeen Hassan Muhammad, Joyce Nakatumba-Nabende, Perez Ogayo, Anuoluwapo Aremu, Catherine Gitau, Derguene, J. Mbaye, Seid Muhie Alabi, Tajuddeen R Yimam, Ignatius 515 Gwadabe, Rubungo Ezeani, Andre Jonathan, Verrah A Mukiibi, Iroro Otiende, Paul Rayson, Mofetoluwa Adeyemi, Gerald Muriuki, E. Anebi, Chiamaka Ijeoma, Chukwuneke, N. Odu, Eric Peter Wairagala, S. Oyerinde, Tobius Clemencia Siro, Saul Bateesa, Deborah Nabagereka, Maurice Katusiime, Ayodele, Mouhamadane Awokoya, Dibora Mboup, Gebrey-525 Henok, Kelechi Tilaye, Nwaike, Degaga, Chantal Amrhein, Rico Sennrich. 2020, On Roman-535, Rosana Ardila, Megan Branson, Kelly Davis, Michael Henretty, Josh Kohler, Reuben Meyer, Alexei Baevski, Wei-Ning Hsu, Alexis Conneau, Tom Brown, Benjamin Mann, Nick Ryder, Jared D Subbiah, Prafulla Kaplan, A. Dhariwal, P. Neelakantan, Girish Shyam, Amanda Sastry, Sandhini Askell, Ariel Agarwal, Herbert-Voss, Gretchen Krueger, T. Henighan, R. Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Ma-teusz Litwin, Scott Gray, B. Chess, Christopher Clark, Sam Berner, Alec McCandlish, Ilya Radford, Sutskever Dario, Amodei, Davis David, Linting Xue, Aditya Barua, Noah Constant, Rami Al-695 Rfou, Sharan Narang, Mihir Kale, Adam Roberts

An overview of Sentiment Analysis Approaches

An overview of Sentiment Analysis Approaches

Shamsuddeen Hassan Muhammad

A Framework for Implementation of E-Classroom System

Khalid Haruna, Baffa, Shamsuddeen Hassan Muhammad, U. Abubakar

Massive Open Online Courses: Awareness, Adoption, Benefits and Challenges in Sub-Saharan Africa

Massive Open Online Courses: Awareness, Adoption, Benefits and Challenges in Sub-Saharan Africa

Shamsuddeen Hassan Muhammad, A. Mustapha, Khalid Haruna

Massive Open Online Courses: A Success of Cloud Computing in Education

Massive Open Online Courses: A Success of Cloud Computing in Education

A. Mustapha, Shamsuddeen Hassan Muhammad, Saheed Abdullahi Salahudeen

OcRI 2016

Prompt, Condition, and Generate Classification of Unsupported Claims with In-Context Learning

Prompt, Condition, and Generate Classification of Unsupported Claims with In-Context Learning

Peter Ebert Christensen, Srishti Yadav, Serge Belongie, Victor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach, Lintang Sutawika, Zaid Alyafeai, Antoine Chaffin, Arnaud Stiegler, Arun Raja, Manan Dey, Saiful Bari, Canwen Xu, Urmish Thakker, Shanya Sharma, Eliza Szczechla, Taewoon Kim, Gunjan Chhablani, Nihal Nayak, Debajyoti Datta, Mike Jonathan Chang, Tian-Jian Jiang, Han Wang, Matteo Manica, Sheng Shen, Zheng-Xin Yong, Harshit Pandey, Rachel Bawden, Thomas Wang, Trishala Neeraj, Jos Rozen, Abheesht Sharma, A. Santilli, Thibault Févry, Jason Alan Fries, Ryan Teehan, Teven Le Scao, Stella Biderman, Leo Gao, Thomas Wolf, Alexander M Rush. 2022, Multi-task, Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, E. Chi, Quoc Le, Denny Zhou. 2023, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ili´c, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luc-cioni, François Yvon, Matthias Gallé, J. Tow, Pawan Sasanka, Thomas Ammanamanchi, Benoît Wang, N. Sagot, Albert Muennighoff, Olatunji Vil-lanova del Moral, Rachel Ruwase, Bawden Stas, A. Bekman, Iz McMillan-Major, Huu Belt-agy, Lucile Nguyen, Samson Saulnier, Pe-dro Tan, Victor Ortiz Suarez, Hugo Sanh, Laurençon Yacine, Julien Jernite, Margaret Launay, Mitchell Colin, Aaron Raffel, Adi Gokaslan, Aitor Simhi, Alham Soroa, Fikri Aji, Amit Alfassy, Anna Rogers, Ariel Kreisberg Nitzav, Chenghao Mou, Chris Emezue, Christopher Klamm, Colin Leong, Daniel Alexander van Strien, D. Adelani, Dragomir R. Radev, E. G. Ponferrada, Efrat Lev-kovizh, Eyal Bar Natan Ethan Kim, F. Toni, Gérard Dupont, Germán Kruszewski, Giada Pistilli, Hady ElSahar, Hamza Benyamina, Hieu Tran, Kimbo Chen, Kyle Lo, Leandro von Werra, Leon Weber, Long Phan, Loubna Ben, Ludovic Tanguy, Manuel Romero Muñoz, Maraim Masoud, María Grandury, Mario Šaško, Max Huang, Maximin Coavoux, Mayank Singh Mike, Minh Chien Vu, M. A. Jauhar, Mustafa Ghaleb, Nishant Subramani, Nora Kassner, Nurulaqilla Khamis, Olivier Nguyen, Omar Espejel, Ona de Gibert, Paulo Villegas, Peter Henderson, Pierre Colombo, Priscilla Amuok, Quentin Lhoest, Rheza Harliman, Rishi Bommasani, Roberto Luis López, Rui Ribeiro, Salomey Osei, S. Pyysalo, Se-bastian Nagel, Shamik Bose, Shamsuddeen Hassan Muhammad, Shayne Longpre, Somaieh Nikpoor, Stanislav Silber-berg, Suhas Pai, S. Zink, Tiago Timponi, Timo Schick, Tristan Thrush, V. Danchev, Vassilina Nikoulina, Veronika Laippala, Violette Lepercq, V. Prabhu, Zeerak Ta-lat, Benjamin Heinzerling, C. Davut, Emre Ta¸sar, Elizabeth Salesky, Sabrina J. Mielke, Wilson Y. Lee, A. Santilli, Debajyoti Datta, Hendrik Strobelt, M. S. Bari, Maged S. Al-Shaibani, Nihal Nayak, Samuel Albanie, Srulik Ben-David, T. Bers, Trishala Neeraj, Deepak Narayanan, Hatim Bourfoune, Jared Casper, Jeff Rasley, Max Ryabinin, Mayank Mishra, Minjia Zhang, Mohammad Shoeybi, Myriam Peyrounette, Liam Hazan, Marine Carpuat, Miruna Clinciu, Na-joung Kim, Newton Cheng, O. Serikov, Omer Antverg, Oskar van der Wal, Rui Zhang, Ruochen Zhang, Sebastian Gehrmann, Shachar Mirkin, Shani Pais, Tatiana Shavrina, Thomas Scialom, Tian Yun, Tomasz Limisiewicz, Verena Rieser, Vitaly Protasov, V. Mikhailov, Yada Pruksachatkun, Yonatan Belinkov, Zachary Bamberger, Zdenvek Kasner, Alice, Carlos Muñoz Saxena, Danish Ferrandis, David Contrac-tor, Davis Lansky, Douwe David, A. KielaDuong, Edward Nguyen, Emi Tan, Ez-inwanne Baylor, Fatima Ozoani, Frankline Onon-iwu Mirza, Habib Rezanejad, H.A. Jones, Indrani Bhat-tacharya, Irene Solaiman, Irina Sedenko, Isar Ne-jadgholi, Jesse Passmore, Joshua Seltzer, Julio Bonis Sanz, L. Dutra, Mairon Samagaio, Maraim El-badri, Margot Mieskes, Marissa Gerchick, Martha Akinlolu, Michael McKenna, Mike Qiu, M. Ghauri, Mykola Burynok, Nafis Abrar, Nazneen Ra-jani, Nour Elkott, N. Fahmy, Olanrewaju Samuel, Ran An, R. Kromann, Ryan Hao, Samira Al-izadeh, Sarmad Shubber, Silas Wang, Sourav Roy, S. Viguier, Thanh Le, Tobi Oyebade, Trieu Le, Yoyo Yang, Abhinav Zach Nguyen, Ramesh Kashyap, Alfredo Palasciano, Alison Callahan, Anima Shukla, Antonio Miranda-Escalada, Ayush Singh, Benjamin Beilharz, Bo Wang, Caio Brito, Chenxi Zhou, Chirag Jain, Chuxin Xu, Clémentine Fourrier, Daniel León Periñán, Daniel Molano, Dian Yu, Enrique Manjava-cas, Fabio Barth, Florian Fuhrimann, Gabriel Altay, Giyaseddin Bayrak, Gully Burns, Helena U. Vrabec, I. Bello, Ishani Dash, Jihyun Kang, John Giorgi, Jonas Golde, J. Posada, Karthik Ranga-sai, Lokesh Sivaraman, Lu Bulchandani, Luisa Liu, Madeleine Shinzato, Maiko Hahn de Bykhovetz, Marc Takeuchi, Maria A Pàmies, Mari-anna Castillo, M. Nezhurina, Matthias Sänger, Samwald Michael, Michael Cullan, Michiel Weinberg, Mina De Wolf, M. Mihaljcic, Moritz Liu, Freidank Myungsun, Natasha Kang, Nathan Seelam, Dahlberg Nicholas, Nikolaus Michio Broad, Pascale Muellner, Patrick Fung, Ramya Haller, Re-nata Chandrasekhar, Robert Eisenberg, Rodrigo Martin, Ros-aline Canalli, Ruisi Su, Samuel Su, Samuel Cahyawijaya, S. GardaShlok, Shubhanshu Deshmukh, Sid Mishra, Simon Ki-blawi, S. Ott, Srishti Sang-aroonsiri, Stefan Kumar, Sushil Schweter, Tanmay Bharati, Laud Théo, Tomoya Gigant, Wojciech Kainuma, Ya-nis Kusa, Labrak, Yash Shailesh, Yash Bajaj, Venkatraman Yifan, Yingxin Xu, Yu Xu, Zhe Xu, Zhongli Tan