I work on machine learning for speech and audio — with a focus on speaker recognition, audiovisual speech processing, voice biometrics, and NLP. I received my PhD from NTUA (2011), held postdoctoral positions at ÉTS Montréal and CRIM (2011–2016), and at the University of Nottingham as a Marie Skłodowska-Curie Fellow (2016–2018). I was elected Associate Professor at AUEB in 2023, where I direct the Information Processing Laboratory and the MSc Programme in AI & Data Science.
My research sits at the intersection of deep learning, speech processing, and biometrics. I am particularly interested in representation learning for audio and speech, with applications ranging from speaker verification to multimodal understanding. Current directions include self-supervised models for speech, parameter-efficient adaptation, and neural diarization.
Speaker embeddings, i-vectors, PLDA, end-to-end verification, voice anti-spoofing, and domain adaptation for speaker identity.
Lip reading, audiovisual fusion, end-to-end ASR, and multimodal models for robust speech perception.
Representation learning from unlabelled speech with HuBERT, Wav2Vec 2.0, WavLM, and downstream adaptation strategies.
Intent classification, task-oriented dialogue, and end-to-end architectures bridging ASR and NLU.
Attentive pooling methods, label smoothing, and cross-corpus generalization for emotion and affect recognition.
End-to-end neural diarization, Perceiver-based attractor models, and overlap-aware segmentation.
I am currently supervising PhD students at AUEB. If you are interested in joining the group, please get in touch by email.
Courses at the Department of Informatics, Athens University of Economics and Business.
Previously also taught "Speech Synthesis and Recognition" and "Dialog Systems" at the Interdisciplinary MSc Programme in Language Technology, National and Kapodistrian University of Athens, in collaboration with the Institute for Language and Speech Processing (ILSP / "Athena" R.C.) — 2022–2024.