A Brief History of ASR: Automatic Speech Recognition
This moment has been a long time coming. The technology behind speech recognition has been in development for over half a century, going through several periods…
AI/ML Engineer · Cairo, Egypt
For ten years I’ve worked on one problem in three costumes: pulling structure out of signal that doesn’t have any. Phonemes out of audio, facts out of documents, answers out of an enterprise’s own tangled data.
Right now that means production LLM systems at an early-stage startup — multi-agent orchestration, retrieval that survives real customer data, and knowledge graphs holding it together.

The through-line
Getting phonemes, dialects, and emotion out of raw audio.
Getting entities, facts, and answers out of documents.
Getting work done across an enterprise’s own messy data.
Selected work
Built and deployed LLM systems for 14+ isolated enterprise customers.
Built Graph RAG and question-answering functionality for Implicit’s Answers product.
Implemented an Arabic speech recognition system for broadcast speech using Kaldi and VariKN.
Research
Toolkit
Writing
This moment has been a long time coming. The technology behind speech recognition has been in development for over half a century, going through several periods…
In this page we will summarize the components of the Arabic ASR project, including the training process and the pipeline, and the pipeline which takes as input a…
In this post, we will summarize the work done on Arabic speech recognition and Arabic dialect identification projects for RedHenLab as part of GSoC 2018. Each of…