Data Engineer, Applied Machine Learning, Text-to-Speech
WellSaid LabsAbout the role
Who We Are: WellSaid Labs
We’re creating Voice for everyone.
At WellSaid Labs, we enable creatives around the globe by putting high-tech, human parity technology into their hands, giving them the ability to add voice-over to any project and iterate with ease. Creative teams use WellSaid Lab’s Voice Studio to create compelling employee training, design unique digital experiences, and narrate audiobooks. We believe deeply in AI for Good, and that technology should be empowering, engaging, and fair to all people.
Who You Are: a Data-Driven Applied ML Engineer, working in Text-to-Speech
The WellSaid Labs Applied Machine Learning Team works with stakeholders to identify, refine, and solve problems at the intersection of machine learning and customer needs. This team understands customer needs through quantitative and qualitative research and they work across WSL teams to understand how machine learning is utilized and how it can be improved.
Improving and maintaining our ML solutions includes creating test datasets and metrics to define and gauge success, working with the ML Platform Team to prioritize model updates, training new models for deployment, coordinating releases, and educating the customer on any new capabilities. The Applied ML Team is consistently testing and reviewing any deployed models.
As a Text-to-Speech Data Engineer on the Applied ML Team at WellSaid Labs, you will be working to regularly improve our text-to-speech service. You will add new datasets; train, deploy, and evaluate new models; and design experiments and algorithms for solving new and creative TTS challenges. You will report your findings to the Applied ML Team who is then actioning any improvements. It would be great for you to be familiar with language construction, voice over and/or audio processing. You may also work closely with our Sr Program Manager: Voice development in establishing data requirements, assessing audio, and generating new script content for voice talent to read.
How You’ll Contribute:
In your day-to-day, you will:
- Work directly with text and audio data: gathering, compiling, and organizing datasets, preparing data for machine training, evaluating results, debugging problematic data
- Train and deploy ML models: incorporate new data, monitor training metrics, debug failing code, deploy models for customer use
- Evaluate and improve ML models: consider causation or correlation between training data and ML predictions, design ML experiments and establish success criteria, gather and evaluate metrics including mean opinion scores, incorporate findings into new experiments
- Additional research projects: interesting data or use cases, alternative services and solutions, internal process improvements, new data evaluation tools, expanding TTS into new languages
Additionally, this role requires you to write and execute code that enables you, and others, to perform each of these tasks. It also requires you to think critically about language, dialect, pronunciation, phonemics, and audio dynamics in order to build the highest quality voices and Studio/API experiences for our customers.
What We’re Looking For
To thrive in this role, you ideally have experience with and a solid understanding of ML concepts and best practices, successfully managed datasets and metrics in a ML capacity, coding experience developing tools that can evaluate data and enable you to establish recommendations based on data analysis results, and experience with software releases to production.
- You have worked in a technical team, managing project expectations and communicating your plans, project statuses, and results frequently and comprehensively
- You have worked with a wide array of data types, building analysis tools and establishing success criteria for evaluating the success of data-driven projects
- You have built and deployed ML models for use by a non-technical audience, clearly communicating usage guidelines and best practices
- You have experience building and documenting new processes, especially in a ML pipeline or similar capacity
- You have a strong understanding of the importance of data preparation for ML training, data visualization and metrics for ML assessment, and analysis of ML results
- You have familiarity with software and feature releases and can work closely with a Product team for exposing ML changes to customers
- [Bonus] You have a curiosity and interest in linguistics and acoustics
- [Bonus] You have an interest in eventually managing an Applied ML Team, strategizing workload, and mentoring contributin
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s