Falcon-Emirati: Building an LLM That Actually Speaks the Dialect
TII's new 7B model scores 84.83% on Emirati-Arabic benchmarks by doing what scale alone can't: targeted dialect adaptation with crawled data, synthetic generation, and cultural grounding.