AI product Open source

Audar-ASR-V1

Audar-ASR-V1 is a family of Arabic-first generative speech-recognition models developed by AudarAI. It treats transcription as audio-conditioned next-token prediction with a language-model decoder, rather than a CTC or transducer objective, using a Whisper-style 128-mel audio encoder and a Qwen3 decoder with a 30-second context. The models cover Modern Standard Arabic, major Arabic dialects, code-switched Arabic-English, English, and 30 languages overall.

View repository Visit site Mentioned in 2 videos ↓

Overview

The Flash tier is intended for real-time, edge, on-device, or offline use, while Turbo targets lower error on difficult dialectal and long-form audio. Both tiers share an architecture and prompt interface and can be run through Transformers, GGUF with llama.cpp, or vLLM; the repository provides model pointers, benchmarks, and inference examples. The published models use the AudarAI Open v1.0 and AudarAI Community v1.0 licenses.

What Audar-ASR-V1 is used for

2 uses taken from transcripts — each links to the moment in the video.

  • A family of Arabic-first generative speech-recognition models and an inference hub with model cards, benchmarks, and inference examples. Its Flash and Turbo tiers target realtime or higher-accuracy transcription and support Arabic dialects, code-switching, and 30 languages overall.

  • A family of Arabic-first generative speech-recognition models from Audar AI. It transcribes Modern Standard Arabic, major Arabic dialects, code-switched Arabic-English, and English, with separate real-time/on-device and higher-accuracy model tiers.

Videos mentioning Audar-ASR-V1

2 in the library.