YOU ARE VIEWING ONE ITEM FROM THE AICRIER FEED

NAMAA releases Cohere Speech Tashkeel 2B

AICrier tracks AI developer news across Product Hunt, GitHub, Hacker News, YouTube, X, arXiv, and more. This page keeps the article you opened front and center while giving you a path into the live feed.

// WHAT AICRIER DOES

7+

TRACKED FEEDS

24/7

SCRAPED FEED

Short summaries, external links, screenshots, relevance scoring, tags, and featured picks for AI builders.

NAMAA releases Cohere Speech Tashkeel 2B
OPEN LINK ↗
// 3h agoMODEL RELEASE

NAMAA releases Cohere Speech Tashkeel 2B

Cohere Speech Tashkeel 2B, developed by NAMAA, is an open-source Arabic speech recognition model that transcribes spoken Arabic into fully diacritized text. The model supports complete vocalization, including harakāt, tanwīn, sukūn, shadda, and grammatical case endings, and has already received community 4-bit quantizations, Ruby bindings, and containerized apps.

// ANALYSIS

Precise diacritization in speech recognition addresses a longstanding gap in Arabic NLP, where unvocalized transcriptions often lose critical grammatical and semantic context.

• Comprehensive diacritization: Transcribing case endings and harakāt directly from audio greatly improves downstream translation, text-to-speech, and linguistic analysis.

• Robust community ecosystem: Immediate availability of GGUF/ONNX 4-bit quantizations and multi-platform Ruby bindings makes deploying the 2B model feasible on edge devices and consumer hardware.

• Focus on under-resourced languages: Demonstrates building open AI tooling tailored to specific linguistic complexities beyond English.

// TAGS
arabic-aisttmodel-releasediacritizationopen-sourcehugging-face

DISCOVERED

3h ago

2026-07-22

PUBLISHED

3h ago

2026-07-22

RELEVANCE

7/ 10

AUTHOR

cohere