
Oddin.gg · Valka.ai
About Valka Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and exper...
About Valka
Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content. Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more.
This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond.
We’re looking for an experienced Aplied Research Scientist for Speech Synthesis for a foundational role to join our new team.
You’ll develop text-to-speech and voice cloning models to create synthetic voices for our avatars that sound like public figures.
We expect you to work with state-of-the-art models and push the limits of what voice cloning and TTS can do. This role requires a solid understanding of speech synthesis, NLP, and deep learning. Experience working with large text and speech datasets is highly desirable.
You’ll build efficient training and deployment pipelines for voice models. Part of your job will be designing validation strategies that compare synthetic speech to real recordings, and creating custom metrics to measure quality.
You’ll also help set up the infrastructure for tracking experiments, making results reproducible, and serving models in production. From training on distributed systems to monitoring deployed models, you’ll be involved in the full machine learning workflow.
About Valka.ai Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content. Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more. This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond.
ABOUT US Founded in the US in 2022 and now based in London, UK, Recraft is an AI tool for professional designers, illustrators, and marketers, setting a new standard for excellence in image generation. We designed a tool that lets creators quickly generate and iterate original images, vector art, illustrations, icons, and 3D graphics with AI. Over 3 million users across 190+ countries have produced hundreds of millions of images using Recraft, and we're just getting started. Join a universe of professional opportunities, develop and support large-scale projects, and shape the future of creativity. We are committed to making Recraft an essential, daily tool for every designer and setting the industry standard. Our mission is to ensure that creators can fully control their creative process with AI, providing them with innovative tools to turn ideas into reality. If you’re passionate about pushing the boundaries of AI, we want you on board! ABOUT THE ROLE As a Junior Machine Learning Engineer at Recraft, you will have the opportunity to work on real-world AI applications, gaining hands-on experience in model development, data collection, evaluation, and production deployment. You will collaborate with research scientists, engineers, and product teams to help enhance Recraft’s AI-driven creative tools. If you are passionate about machine learning, deep learning, and AI-driven applications, this is a great opportunity to learn and contribute to impactful projects. KEY RESPONSIBILITIES * Assist in training, testing, and evaluating machine learning models for real-world applications. * Support data collection and processing for model development. * Conduct experiments and model evaluations, helping improve accuracy and efficiency. * Develop and train large-scale generative models, pushing the boundaries of AI capabilities. * Work closely with ML engineers and researchers to implement AI techniques into production workflows. * Stay updated with the latest trends in AI and deep learning, contributing fresh ideas to the team. QUALIFICATIONS * Currently pursuing a Bachelor’s, Master’s, or PhD in Computer Science, Machine Learning, AI, or a related field. * Reliable Python coding skills. * Knowledge and understanding of foundational deep learning concepts, possibly in application to computer vision, NLP, or speech synthesis or recognition. * Experience with data preprocessing and model evaluation. * Familiarity with MLOps tools is a plus. * Strong analytical skills and ability to work in a collaborative, fast-paced environment. * English: B2+ (Upper-Intermediate or above), written and spoken. WHAT WE OFFER * Real ML work from day one — not toy tasks, but production systems that solve actual problems. * Close mentorship from engineers and researchers who've built and shipped AI at scale. * A front-row seat to how AI products grow: you'll see the full picture, not just your corner of it. * A team that debates ideas, moves fast, and genuinely enjoys what it builds. * Full-time, on-site in London — because the best breakthroughs still happen in the same room. * Skilled Worker visa sponsorship available for the right candidate.
Mirelo AI is building the next generation of creative tools by generating realistic sound, speech and music from video. We develop cutting-edge foundational generative AI models that "unmute" silent video content and create custom, hyper-realistic audio for gaming, video platforms, and creators. Our technology empowers global storytellers to transform their content. We recently closed a $41 million Seed round co-led by Andreessen Horowitz and Index Ventures with participation from Atlantic, and are rapidly expanding across Product, Engineering, Go-to-Market, and Growth. About the Role At Mirelo, you’ll work at the centre of how we build the next generation of multimodal video-to-audio models. This role is deeply hands-on and research-heavy: with a great H100/200-per-engineer ratio you explore and build new multimodal models and push the boundaries of what’s possible in music, sound, and speech generation. You’ll collaborate closely across research and engineering, run focused ablations, and translate experimental results into clear next steps for the team. From data curation to deployment, you’ll help shape the full lifecycle of the models that power our products and partnerships. KEY RESPONSIBILITIES * Design, implement and train large-scale multimodal generative models for audio generation (diffusion and/or autoregressive models). * Explore new modeling ideas for audio generation (music, sound, speech) while taking inspiration from the language and image domains. * Develop and experiment with post-training for new capabilities (fine-grained control, in/out-painting, editing, …) * Conduct rigorous ablation studies, get actionable insights and communicate results to the team to discuss new research directions. * Contribute hands-on to all stages of model development including data curation, experimentation, evaluation, and deployment. IDEAL CANDIDATE PROFILE * Hands-on experience in training large-scale generative models in a fast-paced research environment. * Deep understanding of cutting-edge methods and ML research in at least one of the domains: image, language, video or audio (specific audio experience not necessary, but nice to have). * Strong proficiency in PyTorch, transformer architectures, and the full ecosystem of modern deep learning. * Solid understanding of distributed training techniques—FSDP, low precision training, model parallelism * Strong track-record in working on generative models (publications in top-tier venues, open-source contributions or applied ML projects). NICE TO HAVE * Proficiency with profiling, debugging, and optimizing single and multi-GPU operations using tools like Nsight or stack trace viewers. * Strong software engineering skills/experience in collaborating on large codebases that go beyond PhD research code. * Experience with generative models for audio (sound, music or speech) and audio codec design. WHY JOIN? * Join at a pivotal moment. We've secured fresh funding and are gaining traction - now is when your contributions can make a real difference to our success. * True ownership from day one. You'll have genuine autonomy and responsibility. Your ideas and work will directly shape our product and company direction. * Competitive compensation and equity. We offer strong packages that ensure you share in the success you help create. * Build for the next generation of creators. Be part of the innovation that will transform how creators work and thrive. We welcome applications from all individuals, regardless of ethnic origin, gender, disability, religion or belief, age, or sexual orientation and identity.