Jobs and Careers
OD
Research Scientist, Real-Time DiT Video Models
OddinUKRemotefull_timeVerifiedPosted 31 Mar 2026
About the role
About Valka
Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content. Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more.
This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond.Key Responsibilities:
Design, develop, and optimize AI video generation models using diffusion techniques, with a focus on maintaining consistency, realism, and style.Develop and implement state-of-the-art algorithms for generating video content.Work closely with other teams on large-scale video datasets, including human motion and gestures, facial expressions, and scene context.Experiment with cutting-edge diffusion architectures for controllable and high-quality video synthesis.Define robust validation strategies and implement custom evaluation metrics comparing synthetic vs. real gameplay.Stay on the bleeding edge of the relevant literature, e.g., CVPR, NeurIPS, ICML, ICCV, and help to align it with our roadmap.
Required Qualifications:
🎓 PhD in Computer Vision, Machine Learning, or a closely related field.👉 Hands-on experience with research of video diffusion models.🤝 Strong research background in image or video synthesis demonstrated by having publications at top tier conferences such as CVPR, NeurIPS, ICCV, Siggraph, and ICML.🚀 Experience with real time video generation is a big plus.💼 Proficiency in Python and ML frameworks such as PyTorch🪨 Solid understanding of neural architectures or paradigms (CNNs, Transformers, diffusion models, autoregressive models etc.).
Join us at Valka to lead a new wave of interactive video content, one where your creativity and technical prowess will help transform entire industries and reimagine how digital content is created, shared, and experienced.
This role offers a unique opportunity to shape the future of interactive video content, where digital humans can engage in meaningful and dynamic interactions with users. If you're a passionate ML expert with a drive to innovate and create immersive experiences, we encourage you to apply.
Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content. Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more.
This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond.Key Responsibilities:
Design, develop, and optimize AI video generation models using diffusion techniques, with a focus on maintaining consistency, realism, and style.Develop and implement state-of-the-art algorithms for generating video content.Work closely with other teams on large-scale video datasets, including human motion and gestures, facial expressions, and scene context.Experiment with cutting-edge diffusion architectures for controllable and high-quality video synthesis.Define robust validation strategies and implement custom evaluation metrics comparing synthetic vs. real gameplay.Stay on the bleeding edge of the relevant literature, e.g., CVPR, NeurIPS, ICML, ICCV, and help to align it with our roadmap.
Required Qualifications:
🎓 PhD in Computer Vision, Machine Learning, or a closely related field.👉 Hands-on experience with research of video diffusion models.🤝 Strong research background in image or video synthesis demonstrated by having publications at top tier conferences such as CVPR, NeurIPS, ICCV, Siggraph, and ICML.🚀 Experience with real time video generation is a big plus.💼 Proficiency in Python and ML frameworks such as PyTorch🪨 Solid understanding of neural architectures or paradigms (CNNs, Transformers, diffusion models, autoregressive models etc.).
Join us at Valka to lead a new wave of interactive video content, one where your creativity and technical prowess will help transform entire industries and reimagine how digital content is created, shared, and experienced.
This role offers a unique opportunity to shape the future of interactive video content, where digital humans can engage in meaningful and dynamic interactions with users. If you're a passionate ML expert with a drive to innovate and create immersive experiences, we encourage you to apply.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s
Similar roles
Applied AI Engineer & Researcher
Speechify
Minneapolis
$200,000/yr
Research Assistant, Language and Development Internship (Open to Current UT Students Only)
The University of Texas at Austin
UT MAIN CAMPUS, United StatesRemote
$2,147,483,647/yr
Research Assistant (Bilingual - English/Spanish)
San Ysidro Health
San Diego