CV
Saeedreza Zouashkiani
Summary
AI Researcher with expertise in Speech Processing, Large Language Models, and Reinforcement Learning. Currently at the Foundation AI Models Lab, Sharif University of Technology.
Education
- M.Sc. in Communication Systems2024-06Sharif University of TechnologyGPA: 3.7/4.0Courses: Deep Learning, Convex Optimization, Adaptive Filters
- B.Sc. in Electrical Engineering2021-09Amirkabir University of TechnologyGPA: 3.2/4.0Courses: Statistical Learning, Linear Algebra
Work Experience
- AI Researcher, Automatic Speech Recognition2024-08 - PresentFoundation AI Models Lab, Sharif University of TechnologyConducted research on ASR architectures, trained top-performing models, and led a project integrating continual dataset growth, achieving state-of-the-art results for Persian ASR.
- AI Consultant2025-05 - PresentBimeh Dot Com, Sepas HoldingDeveloped automated pipelines for document understanding, researched RL applications in finance, and built an OCR pipeline for car insurance card processing.
- AI Researcher, Text-to-Speech2025-03 - PresentFoundation AI Models Lab, Sharif University of TechnologyDeveloped a G2P pipeline for Persian, compiled a high-quality audiobook corpus, and trained the first bilingual Persian-English zero-shot TTS model.
- AI Researcher, Voice Conversion2025-03 - PresentFoundation AI Models Lab, Sharif University of TechnologyLeading research on voice conversion, including emotion transferability, and fine-tuning a model on 1,000 hours of Persian audio data.
- AI Researcher, Speech Data Creation Pipeline2024-08 - 2025-07Foundation AI Models Lab, Sharif University of TechnologyCollected and processed over 11.8k hours of audio data, created a scalable validation pipeline, and produced large-scale datasets for ASR, TTS, and VC.
- Graduate Researcher, MSc Thesis2021-09 - 2024-06Sharif University of TechnologyProposed a hybrid meta-reinforcement learning framework for dynamic edge offloading, reducing trainable parameters and leveraging deep learning for scalable edge orchestration.
Skills
Programming
- Python
- Vibe Coding
- MATLAB
- Bash
AI Frameworks
- PyTorch
- TensorFlow
- Scikit-learn
- OpenCV
- Hugging Face
AI Domains
- Speech Processing
- Natural Language Processing
- Computer Vision
- Reinforcement Learning
- Prompt Engineering
Simulation & Tools
- Git
- Docker
- Linux (Ubuntu)
Publications
- PersianVox: A Prosody-Aware Approach for Speech Dataset Generation from In-the-Wild Data2026arXiv:2609.19324Zouashkiani, S., Khalesi, S., Soleimani Roudi, S., Amini, S., and Ghaemmaghami, S.
- A Survey on Non-Intrusive ASR Refinement: From Output-Level Correction to Full-Model Distillation2025arXiv:2508.07285Peyghan, M. R., Rajabi, F., Soleimani Roudi, S., Zouashkiani, S., Amini, S., and Ghaemmaghami, S.
Teaching
- Deep Learning2024Sharif University of TechnologyRole: Head Teaching Assistant
- Introduction to Machine Learning2023Sharif University of TechnologyRole: Teaching Assistant
- Machine Learning2023Sharif University of TechnologyRole: Course Coordinator
- Communication Systems2023Sharif University of TechnologyRole: Teaching Assistant
- Adaptive Filter2022Sharif University of TechnologyRole: Teaching Assistant
- Linear Algebra2021Amirkabir University of TechnologyRole: Guest Lecturer
Portfolio
- Persian ASR, Translation & Summarization App2025GradioDeveloped a Gradio app for Persian ASR, with support for multilingual translation and summarization of transcriptions using Gemini API.
- ParsNorm: Advanced Persian Text Normalization2024GithubDeveloped a comprehensive text normalization pipeline and a unique English-to-Persian transliteration tool.
- Pin Insulator Instance Segmentation with YOLOv82023GithubTrained a YOLOv8 segmentation model on the CLIPD dataset, achieving 96.41% mAP50.
Languages
- PersianNative
- EnglishFluent
Interests
- Speech ProcessingAutomatic Speech Recognition (ASR), Text-to-Speech (TTS), Voice Conversion (VC)
- Large Language ModelsRetrieval-Augmented Generation (RAG), Agentic Frameworks
- Reinforcement LearningOptimization Techniques, Exploration Strategies