
Oddin.gg · Valka.ai
About Valka.ai Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create an...
About Valka.ai
Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content.
Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more.
This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond.
About Valka Valka, a visionary spin-off from the Realms Group (the parent company of Oddin.gg), is on a mission to revolutionize the way people create and experience digital content. Our team believes that content shouldn’t just be consumed; it should be co-created in real time, blurring the lines between imagination and reality. By harnessing the power of cutting-edge AI, we aim to build an interactive human-digital platform where virtual characters respond dynamically to each user’s voice, text, gestures, and more. This is your chance to join a diverse group of innovators who are driven to redefine what’s possible in generative content. Together, we’re changing the paradigm from passive viewing to active participation, unlocking new creative frontiers across gaming, entertainment, education, and beyond. Position Intro: We’re looking for an experienced Aplied Research Scientist for Speech Synthesis for a foundational role to join our new team. You’ll develop text-to-speech and voice cloning models to create synthetic voices for our avatars that sound like public figures. We expect you to work with state-of-the-art models and push the limits of what voice cloning and TTS can do. This role requires a solid understanding of speech synthesis, NLP, and deep learning. Experience working with large text and speech datasets is highly desirable. You’ll build efficient training and deployment pipelines for voice models. Part of your job will be designing validation strategies that compare synthetic speech to real recordings, and creating custom metrics to measure quality. You’ll also help set up the infrastructure for tracking experiments, making results reproducible, and serving models in production. From training on distributed systems to monitoring deployed models, you’ll be involved in the full machine learning workflow.
SENIOR DATA SCIENTIST – STEALTH AI / DATA STARTUP ABOUT NEON Many large companies make billions each year by monetizing Americans’ personal data. At Neon, we’re finally cutting consumers in on the deal. Neon allows our users to make hundreds (or even thousands) of dollars per year by securely selling their anonymized data. We recently raised over $25 million from investors including Lightspeed, Upper90, and Upfront Ventures. About the Role We’re building something big at the intersection of media, consumer data, and AI. We’re seeking a Senior Data Scientist to define and drive the analytical and experimental work that connects our data systems to real business and customer outcomes. You’ll be a strategic partner across product and engineering, helping shape decisions around experimentation, tooling, and insights that matter most to our users. WHAT YOU’LL BE WORKING ON * Own end-to-end experimentation and analytical workflows that help the business learn fast and make confident decisions rooted in data. * Work closely with product, data engineering, and business stakeholders to start with customer and business goals, and work backwards to design the right experiments, features, metrics, and models. * Focus initially on optimizing our data pipeline for media data (audio and text) — helping improve throughput, quality, and utility of data as it flows through our systems. * Contribute to the evolution of our data infrastructure to support more advanced capabilities — including eventual creation of novel datasets and internal modeling workflows (e.g., for voice or multimodal data). * Design, analyze, and interpret experiments across tools/services/technologies to identify what works best — producing actionable insights that influence product direction, operations, and future research. * Translate complex analytical outcomes into clear business insights and product recommendations, occasionally contributing to broader thought leadership and research narratives. WHAT WE’RE LOOKING FOR * Strong analytical background (e.g., statistics, experimentation, causal inference) with several years of real-world experience in data science or research-oriented roles. * Comfort operating independently and pragmatically in a startup environment, balancing speed and rigor. * Excellent communicator who can bridge technical and business audiences — asking the right questions, framing hypotheses, and presenting findings effectively. * Demonstrated ability to guide experimentation from design through interpretation, and connect results back to product and business levers. NICE TO HAVE * Experience with media-centric data (audio, video, text) and related tooling or models. * Familiarity with machine learning workflows, model evaluation, and production experimentation frameworks. * Additional experience with OOP languages, particularly JavaScript * Bonus points if you’ve worked closely with voice AI, speech, or related media AI labs or contributed to their research efforts. WHAT WE OFFER * Competitive compensation with equity upside * Opportunity to shape direction of analytics and modeling at an early stage * A collaborative team environment and a fun, can-do attitude! :)
SNAPSHOT We are seeking a highly motivated Research Engineer (L5) with a strong background in multi-modal modelling for humans and a focus on speech & audio/visual to join the effort within Google DeepMind's Frontier AI unit. This role is pivotal in developing foundational multimodal AI capabilities to understand, generate, and protect human likeness. As a key contributor, you will design and implement cutting-edge models and frameworks, pushing the boundaries of AI to enable foundational capabilities for human-centric understanding and generation. This is a unique opportunity to contribute to impactful research and advance Google DeepMind's mission towards Artificial General Intelligence (AGI). ABOUT US Artificial Intelligence could be one of humanity’s most useful inventions. At Google DeepMind, we’re a team of scientists, engineers, machine learning experts and more, working together to advance the state of the art in artificial intelligence and ultimately achieve Artificial General Intelligence. We use our technologies for widespread public benefit and scientific discovery, and collaborate with others on critical challenges, ensuring safety and ethics are the highest priority. The effort is a part of Google DeepMind's Frontier AI unit. The team aims to build holistic representation encompassing a full spectrum of human understanding. We develop systems to provide perception skills critical for person-centric applications, which is crucial for enabling AI to interact naturally & seamlessly, depict humans accurately & responsibly in generative AI, and build trustworthy & resilient systems that can detect and prevent misuse like deepfakes and impersonation. THE ROLE You will drive outcomes for critical technical components aimed at advancing our capabilities in multimodal human understanding. You will play a critical role in developing and deploying models that can provide accurate human understanding across multiple modalities (e.g., visual appearance, voice, dynamics, etc), while also building robust defenses against sophisticated AI-driven manipulation and impersonation. This role involves tackling complex, ambiguous problems with no obvious "best" solution, requiring independent judgment and a proactive approach to exploring multiple technical avenues. You will be instrumental in shaping the technical direction for core components of the effort. Your contribution will lead to key breakthrough and impactful landings within GDM and across Google products, ensuring our technologies are both groundbreaking and responsibly deployed. KEY RESPONSIBILITIES * Advance multimodal human representations & understanding : Research and implement novel models and other multimodal techniques for a more holistic understanding of humans across visual, audio, and textual data. * Conduct applied research: Conduct experimental research cycles from hypothesis to deployment. * Drive technical projects: Take ownership of substantial technical projects within the effort, from ideation and design to implementation and evaluation, often involving cross-functional collaboration. * Contribute to Infrastructure: Inform and contribute to the development of scalable and efficient research infrastructure for multimodal human understanding models and datasets. * Design and execute strategies for tuning and adapting VLMs and other foundation models for specific tasks ABOUT YOU In order to set you up for success as a Research Engineer at Google DeepMind, we look for the following skills and experience: Requirements: * PhD degree in Computer Science, Machine Learning, or a related technical field with 3+ years of relevant experience. * Experience in developing machine learning models, such as audio & speech-visual models. * Experience in working with and tuning large-scale vision language models. * Strong programming skills in Python and experience with at least one major deep learning framework (e.g., JAX) * Experience conducting independent research and development, including experimental design, implementation, and analysis. In addition, the following would be an advantage: * Experience with Generative AI techniques and architectures. * Familiarity with Reinforcement Learning or alignment methods. * A track record of publications in top-tier AI/ML conferences (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV). * Experience with multimodal learning, integrating information from different data types (e.g., vision, audio, text). * Understanding of privacy-preserving machine learning or responsible AI practices. The US base salary range for this full-time position is between 174,000 USD - 252,000 USD + bonus + equity + benefits. Your recruiter can share more about the specific salary range for your targeted location during the hiring process. Note: In the event your application is successful and an offer of employment is made to you, any offer of employment will be conditional on the results of a background check, performed by a third party acting on our behalf. For more information on how we handle your data, please see our Applicant and Candidate Privacy Policy. At Google DeepMind, we value diversity of experience, knowledge, backgrounds and perspectives and harness these qualities to create extraordinary impact. We are committed to equal employment opportunities regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, pregnancy, or related condition (including breastfeeding) or any other basis as protected by applicable law. If you have a disability or additional need that requires accommodation, please do not hesitate to let us know.