Apple speech researcher, co-author of dMel speech tokenization
Machine Learning Researcher · current
Zakaria's dMel work showed that a simple, honest representation can beat elaborate speech tokenizers, a result that helps every speech-language model that follows. His spatial audio embeddings research points at genuinely new listening experiences.
We built this from your public work because we think it deserves celebrating. You did not ask us to, so the only fair thing is that you decide what happens to it. Claim it and it is yours to edit. Ask us to change something and we will. Ask us to take it down and it is gone within 72 hours — free, no account needed, and nobody will try to talk you out of it.