Egyptian Arabic Energy Domain Speech Dataset
A single-turn energy domain speech dataset based on the Egyptian Arabic dialect.
Use verified, licensed data with confidence. You can download right away or check the data through inquiry.
A single-turn energy domain speech dataset based on the Egyptian Arabic dialect.
A single-turn public administration domain speech dataset based on Modern Standard Arabic (MSA).
A single-turn public administration domain speech dataset based on the Gulf Arabic dialect.
A single-turn public administration domain speech dataset based on the North African Arabic (Maghrebi Arabic) dialect.
A single-turn public administration domain speech dataset based on the Egyptian Arabic dialect.
A multimodal Arabic image dataset combining image captions and preference QA for general and construction domains, reviewed and annotated by construction and engineering domain experts.
A multilingual multi-turn chat text dataset built directly by professional annotators, including speaker information, proper nouns, chat slang, typo tags, and natural conversational turns.
A multilingual domain-specific parallel translation corpus dataset built from Korean-to-English, Korean-to-Arabic, and Korean-to-Chinese translations specialized in military, finance, and legal domains, curated by expert annotators.
A Korean-based Arena benchmark text dataset built for evaluating AI model performance in general-domain comparison tasks.
A Korean-based BFCL 3v benchmark text dataset built for evaluating function-calling capabilities of AI models.