SonicJobs Logo
Left arrow iconBack to search

Research Engineer, Text-To-Speach

Oddin
Posted 6 months ago, valid for 21 days
Location

Malabar, Brevard County 32950, FL

Salary

Competitive

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • Valka.ai, a spin-off from the Realms Group, is seeking a candidate to research and train state-of-the-art text-to-speech (TTS) models for entertainment and education applications.
  • The role involves maintaining the best TTS model, collaborating with a team of researchers, and ensuring smooth deployment with product engineering.
  • Candidates should have experience working with machine learning models in production, proficiency in Python and key libraries like PyTorch, and familiarity with TTS and voice cloning models.
  • Nice-to-have skills include knowledge of transformers, GANs, and contributions to open-source AI tools, as well as familiarity with cloud providers like AWS.
  • The position requires a minimum of 3 years of experience, and the salary is competitive based on experience.

About Valka.ai


Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content.


Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more.


This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond.


\n


What you will be doing
  • Our goal is to research and train fast and high-quality SOTA TTS models for realistic and emotional voice generation for entertainment and education applications.

  • You will be in charge of maintaining the “current best” TTS model we have - assembling the results of the best experiments into one speech production system, and evaluating it in terms of performance and quality.

  • You will be in immediate collaboration with a our TTS researcher team, and close cooperating with product engineering and platform roles to ensure smooth deployment


Skills you need
  • Experience working with ML modes in production (model formats, quantization, deployment, hardware requirements, model logging and tracking…)
  • Proficiency in Python and key libraries (e.g., PyTorch, Hugging Face Transformers).

Nice to have 

  • Experience with training text-to-speech / voice cloning models, understanding of human speech and audio processing.
  • Knowledge of transformers, diffusion models, GANs.
  • Familiarity with modern speech synthesis models (GPT-based, flow matching… such as Vevo, StyleTTS, IndexTTS, Maskgct etc.).
  • Contributions to open-source AI tools
  • Familiarity with AWS / other cloud providers


\n



Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.