نادي ٢٠٢٦ NADI 2026
The Seventh Edition

نادي ٢٠٢٦

NADI 2026

Robust & Mixed-Dialect Arabic Speech

A shared task on five speech-processing challenges for the Arab world’s living dialects , including recognition, identification, synthesis, translation, and spoken understanding.

Tasks
5 main
Dialects
11 countries
Test release
July 20, 2026
Venue
ArabicNLP 2026
نظرة عامَّة
Overview

About the shared task

Arabic speech systems still struggle: Many current systems perform reasonably well on Modern Standard Arabic (MSA) and clean speech from dominant dialects, but degrade under realistic conditions such as background noise, low-bandwidth audio, mixed dialects, and sub-country regional variation.

Building on NADI 2025’s spoken dialect work, NADI 2026 broadens the focus with three main speech-focused families of tasks: robust dialectal ASR for real-world conditions, spoken dialect identification under cross-domain conditions, and dialectal Arabic text-to-speech as a new generative component — alongside speech translation and spoken language understanding.

Together, these tasks benchmark discriminative and generative Arabic speech systems to advance robust, inclusive technologies reflecting the Arab world’s linguistic diversity.

المسارات
Shared Task Subtasks

Five complementary tasks

Each task ships with new blind test data. Baselines, evaluation scripts, and submission links released with the data on June 16, 2026.

الموارد
Resources

Datasets, notebooks, and platform links

This table summarizes the main resources for each task, including notebooks, training data, development data, and the relevant platform or submission path. Task 3 TTS follows a different evaluation and submission flow, so please check its submission instructions button.

Task Baselines Training data Development data Test data Evaluation
Notebook #hours #utterances Dataset #hours #utterances Dataset Dataset Platform
🏆 لوحات النَّتائج
Leaderboards

Top-performing systems

Task 2

Spoken Dialect ID

Accuracy ↑
  1. 1 ThakaThaka, Advanced AI & Information Technology 56.15
  2. 2 SalesteqSalesteq 54.66
  3. 3 LynxMohammed VI Polytechnic University 53.53
Higher is better14 submissions
Task 3

Dialectal TTS

UTMOS ↑
  1. 1 oddadmix 2.72
Higher is better1 submission
Task 5.1

SLU Intent Classification

Weighted F1 ↑
  1. 1 LynxMohammed VI Polytechnic University 76.86
  2. 2 CIS-OpenSLUNile University 71.54
  3. 3 ahmednezarGlobal Brands Group 69.92
Higher is better7 submissions
Task 5.2

SLU Slot Filling

CoER ↓
  1. 1 AslemaQCRI 59.53
  2. 2 ZilaHacettepe University 72.78
Lower is better2 submissions

Please contact the shared task chairs for any issues with affiliation information, or other issues.

المواعيد المهمَّة
Important Dates

Key milestones

Combined NADI shared-task milestones and ArabicNLP 2026 conference deadlines. Source: arabicnlp2026.sigarab.org.

DateMilestoneStatus
المشاركة
Participation Guidelines

How to participate

1

Register your team

Fill out the registration form to receive access to training and development data, baseline systems, and submission links.

Registration form →
2

Build & submit

Submit via CodaBench for ASR & SID where suitable, and Hugging Face Spaces for large TTS audio submissions. SID hidden test runs through a private platform.

View baselines →
3

Write your paper

Submit a system description paper by August 22. Document external data, pretrained models, preprocessing, and decoding settings clearly.

Paper guidelines →

For questions contact the organizers at nadisharedtask@gmail.com.

إرشادات كتابة الورقة البحثية
System Paper Guidelines

Instructions for writing the system paper

Purpose

What is the paper for?

The system description paper should let another researcher:

Verify what the system does and how it has been trained.
Reimplement the system to reproduce the results.
Understand the system’s strengths and weaknesses.
Conference ArabicNLP 2026
Page limit 4 pages of content
Review Not double-blind
Submission OpenReview — TBA
Format and submission

What format should the paper use?

The paper will be included in the The Fourth Arabic Natural Language Processing Conference (ArabicNLP 2026) proceedings. Please familiarize yourself with the general shared task paper requirements for the conference.

The paper is expected to be up to 4 pages of content, plus unlimited references and appendices; final versions of the paper will be given one additional page of content (up to 5 pages) so that reviewers’ comments can be taken into account. Please, note that the review process is not double-blind, so anonymity is not required.

Paper submissions must use the official ACL style templates, which are available here (Latex). Please follow the paper formatting guidelines, general to "*ACL" conferences available here.

Important LaTeX note

note: use the acl_latex.tex template but do not use acl_lua_latex.tex due to poorer Arabic language support in LuaLaTex.

Authors may not modify these style files or use templates designed for other conferences.

Submission Website: submissions should be done via openreview (TBA)
Recommended organization

How should the papers be structured?

A common structure for system description papers is:

01

Title

Title should be as follows: your_team_name at NADI 2026 shared task: your_own_title. For example, UBC at NADI 2026 shared task: Multitask learning for Arabic Dialect Identification

02

Abstract

four/five sentences highlighting your approach and key results.

03

Introduction

¾ a page expanding on the abstract mentioning key background such as why the task is challenging for current modeling techniques and why your approach is interesting/novel.

04

Data

review of the data you used to train your system. Be sure to mention the size of the training, validation and test sets that you’ve used, and the label distributions, as well as any tools you used for preprocessing data.

05

System

a detailed description of how the systems were built and trained. If you’re using a neural network, were there pre-trained embeddings, how was the model trained, what hyperparameters were chosen and experimented with? How long did the model take to train, and on what infrastructure? Linking to source code is valuable here as well, but the description should be able to stand alone as a full description of how to reimplement the system. While other paper styles include background as a separate section, it’s fine to simply include citations to similar systems which inspired your work as you describe your system.

06

Results

a description of the key results of the paper. If you have done extra error analysis into what types of errors the system makes, this is extremely valuable for the reader. Unofficial results from after the submission deadline can be very useful as well.

07

Discussion

general discussion of the task and your system. Description of characteristic errors and their frequency over a sample of development data. What would you do if you had another 3 months to work on it?

08

Conclusion

a restatement of the introduction, highlighting what was learned about the task and how to model it.

Citation

Please use this bibtex entry to cite the NADI-2026 shared task overview paper:

@inproceedings{Sullivan-etal-2026-nadi,
    title = "{NADI-2026: The Second Multidialectal {A}rabic Speech Processing Shared Task}",
    author = "Sullivan, Peter  and
      Talafha, Bashar  and
      Ashraf, Ahmed  and
      Bougares, Fethi  and
      Elleuch, Haroun  and
      Zhang, Chiyu  and
      Elmadany, AbdelRahim  and
      Mohamed, Youssef  and
      Mdhaffar, Salima  and
      Est{\`e}ve, Yannick  and
      Elhoseiny, Mohamed  and
      Luqman, Hamzah  and
      Habash, Nizar  and
      Abdul-Mageed, Muhammad",
    booktitle = "Proceedings of the Fourth Arabic Natural Language Processing Conference (ArabicNLP 2026)",
    year = "2026",
    address = "Budapest, Hungary",
    publisher = "Association for Computational Linguistics",
}
Prior citations

Also, please use these entries to cite prior NADI papers:

NADI-2020

@inproceedings{abdul-mageed-etal-2020-nadi,
    title = "{NADI} 2020: The First Nuanced {A}rabic Dialect Identification Shared Task",
    author = "Abdul-Mageed, Muhammad  and
      Zhang, Chiyu  and
      Bouamor, Houda  and
      Habash, Nizar",
    booktitle = "Proceedings of the Fifth Arabic Natural Language Processing Workshop",
    month = dec,
    year = "2020",
    address = "Barcelona, Spain (Online)",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2020.wanlp-1.9",
    pages = "97--110",
}

NADI-2021

@inproceedings{abdul-mageed-etal-2021-nadi,
    title = "{NADI} 2021: The Second Nuanced {A}rabic Dialect Identification Shared Task",
    author = "Abdul-Mageed, Muhammad  and
      Zhang, Chiyu  and
      Elmadany, AbdelRahim  and
      Bouamor, Houda  and
      Habash, Nizar",
    booktitle = "Proceedings of the Sixth Arabic Natural Language Processing Workshop",
    month = apr,
    year = "2021",
    address = "Kyiv, Ukraine (Virtual)",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2021.wanlp-1.28",
    pages = "244--259",
}

NADI-2022

@inproceedings{abdul-mageed-etal-2022-nadi,
    title = "{NADI} 2022: The Third Nuanced {A}rabic Dialect Identification Shared Task",
    author = "Abdul-Mageed, Muhammad  and
      Zhang, Chiyu  and
      Elmadany, AbdelRahim  and
      Bouamor, Houda  and
      Habash, Nizar",
    booktitle = "Proceedings of the The Seventh Arabic Natural Language Processing Workshop (WANLP)",
    month = dec,
    year = "2022",
    address = "Abu Dhabi, United Arab Emirates (Hybrid)",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2022.wanlp-1.9",
    pages = "85--97",
}

NADI-2023

@inproceedings{abdul-mageed-etal-2023-nadi,
    title = "{NADI} 2023: The Fourth Nuanced {A}rabic Dialect Identification Shared Task",
    author = "Abdul-Mageed, Muhammad  and
      Elmadany, AbdelRahim  and
      Zhang, Chiyu  and
      Nagoudi, El Moatez Billah  and
      Bouamor, Houda  and
      Habash, Nizar",
    editor = "Sawaf, Hassan  and
      El-Beltagy, Samhaa  and
      Zaghouani, Wajdi  and
      Magdy, Walid  and
      Abdelali, Ahmed  and
      Tomeh, Nadi  and
      Abu Farha, Ibrahim  and
      Habash, Nizar  and
      Khalifa, Salam  and
      Keleg, Amr  and
      Haddad, Hatem  and
      Zitouni, Imed  and
      Mrini, Khalil  and
      Almatham, Rawan",
    booktitle = "Proceedings of ArabicNLP 2023",
    month = dec,
    year = "2023",
    address = "Singapore (Hybrid)",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2023.arabicnlp-1.62",
    doi = "10.18653/v1/2023.arabicnlp-1.62",
    pages = "600--613",
}

NADI-2024

@inproceedings{abdul-mageed-etal-2024-nadi,
    title = "{NADI} 2024: The Fifth Nuanced {A}rabic Dialect Identification Shared Task",
    author = "Abdul-Mageed, Muhammad  and
      Keleg, Amr  and
      Elmadany, AbdelRahim  and
      Zhang, Chiyu  and
      Hamed, Injy  and
      Magdy, Walid  and
      Bouamor, Houda  and
      Habash, Nizar",
    editor = "Habash, Nizar  and
      Bouamor, Houda  and
      Eskander, Ramy  and
      Tomeh, Nadi  and
      Abu Farha, Ibrahim  and
      Abdelali, Ahmed  and
      Touileb, Samia  and
      Hamed, Injy  and
      Onaizan, Yaser  and
      Alhafni, Bashar  and
      Antoun, Wissam  and
      Khalifa, Salam  and
      Haddad, Hatem  and
      Zitouni, Imed  and
      AlKhamissi, Badr  and
      Almatham, Rawan  and
      Mrini, Khalil",
    booktitle = "Proceedings of the Second Arabic Natural Language Processing Conference",
    month = aug,
    year = "2024",
    address = "Bangkok, Thailand",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2024.arabicnlp-1.79/",
    doi = "10.18653/v1/2024.arabicnlp-1.79",
    pages = "709--728",
}

NADI-2025

@inproceedings{talafha-etal-2025-nadi,
    title = "{NADI} 2025: The First Multidialectal {A}rabic Speech Processing Shared Task",
    author = "Talafha, Bashar  and
      Toyin, Hawau Olamide  and
      Sullivan, Peter  and
      Elmadany, AbdelRahim A.  and
      Juma, Abdurrahman  and
      Djanibekov, Amirbek  and
      Zhang, Chiyu  and
      Alshehhi, Hamad  and
      Aldarmaki, Hanan  and
      Jarrar, Mustafa  and
      Habash, Nizar  and
      Abdul-Mageed, Muhammad",
    editor = "Darwish, Kareem  and
      Ali, Ahmed  and
      Abu Farha, Ibrahim  and
      Touileb, Samia  and
      Zitouni, Imed  and
      Abdelali, Ahmed  and
      Al-Ghamdi, Sharefah  and
      Alkhereyf, Sakhar  and
      Zaghouani, Wajdi  and
      Khalifa, Salam  and
      AlKhamissi, Badr  and
      Almatham, Rawan  and
      Hamed, Injy  and
      Alyafeai, Zaid  and
      Alowisheq, Areeb  and
      Inoue, Go  and
      Mrini, Khalil  and
      Alshammari, Waad",
    booktitle = "Proceedings of The Third Arabic Natural Language Processing Conference: Shared Tasks",
    month = nov,
    year = "2025",
    address = "Suzhou, China",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2025.arabicnlp-sharedtasks.99/",
    doi = "10.18653/v1/2025.arabicnlp-sharedtasks.99",
    pages = "720--733",
    ISBN = "979-8-89176-356-2",
}
المُنَظِّمون
Organizing Committee

Task organizers

انضمَّ إلى نادي ٢٠٢٦

Join the seventh edition

Registration opens May 16, 2026. Training and development data, baseline systems, and evaluation scripts land June 16. Blind test data ships July 20.