Prosody-Conditioned Speech & Language Modeling
Focus: Speech representation learning for pathological speech processing
Machine Learning Researcher & Engineer
Open to research collaborations
I'm passionate about bridging the gap between cutting-edge research and practical deployment. My research focuses on representation learning and foundation models for scientific and biomedical AI. Alongside academic research, I have built production machine learning systems in industry, shaping my interest in methods that are not only novel but also scalable, reproducible, and clinically meaningful.
With hands-on experience as an ML Engineer at Causalmind, I've built production systems using causal discovery and LLMs, deploying models on the cloud with Docker and managing complex data pipelines. This industry experience informs my research approach: I prioritize not just theoretical advances, but solutions that are interpretable, scalable, and clinically applicable. I believe the future of AI lies in creating foundation models that generalize across domains while remaining robust enough for healthcare and scientific applications.
Focus: Speech representation learning for pathological speech processing
Focus: Hierarchical & multi-scale representation learning for clinical assessment
This project explores representation learning and multimodal alignment for audio foundation models with applications in scientific and healthcare domains.
Developing hierarchical federated learning systems that enable efficient collaboration across distributed institutions while maintaining strict privacy constraints through encryption and differential privacy techniques.
Designing comprehensive frameworks for AI literacy and competence development in higher education, addressing how students and educators can effectively understand, evaluate, and responsibly deploy AI systems.
Coming soon
Coming soon
Zhejiang University
July 2026 - Present
Zhejiang University of Science and Technology
January 2026 - Present
Causalmind | Causal Discovery & AI Research
September 2025 - December 2025
Zhejiang University of Science and Technology
Major: Data Science and Big Data Technology
Expected Graduation: 2027
Advanced research projects on audio, speech, and multimodal AI
Building audio foundation model for scientific and clinical applications.
Speech-based machine learning models for neurodegenerative disease detection using hierarchical representation learning.
Speech representation learning for pathological speech processing with prosody conditioning.
Deep learning for scientific discovery and healthcare applications
A deep learning project using CNNs and transfer learning to classify brain CT scans, identifying tumor types or normal brains, helping accelerate and improve diagnostic accuracy.
Quantitative Structure-Activity Relationship modeling for drug discovery with ADMET property prediction and therapeutic target evaluation.
Production ML systems for real-world applications
Machine learning models to predict credit card approval based on applicant data such as income, credit score, and employment history, helping automate decisions and improve accuracy.
Implemented Linear Regression model to forecast house prices, emphasizing data preparation, and feature evaluation and model performance metrics.
Kelas.com
September 2025
Certificate ID: CERT-13B86355
Verify CertificateInterested in discussing opportunities, collaborations, or just want to chat about data science and AI? I'd love to hear from you!
I typically respond within 24 hours
Schedule a Meetingme@oliviastefany.dev | (+86)18668233076
Currently based in China | Open to remote opportunities