About this AI Researcher - Video Generation role at D-ID
Join D-ID’s research team to develop the next generation of multimodal video diffusion models and bring cutting-edge research into real-time products used at global scale.
About the Team
Our research team is small, fast-moving, and highly impactful. We develop multimodal generative models for real-time video creation, turning cutting-edge research into scalable products.
We are looking for researchers with deep expertise in generative video to join a team at the core of D-ID’s technology and help shape the future of generative AI.
Why Work at D-ID?
- Work on cutting-edge video-generation technology at the forefront of AI.
- See your research become part of products used at global scale.
- Own meaningful research and engineering challenges as part of a small, high-impact team.
- Help solve complex problems in real-time video generation.
- Build technology that brings the next generation of AI assistants to life.
Role & Responsibilities
- Research, develop, and evaluate generative AI models for video generation.
- Design experiments and evaluation methods covering visual quality, temporal consistency, identity preservation, and performance.
- Explore, validate, and implement innovative research ideas.
- Own the full lifecycle from research and experimentation to production deployment.
- Build models and systems that perform reliably at scale and in real time.
- Collaborate closely with research, engineering, and product teams.
You are
- A curious researcher and dedicated self-learner.
- Independent, proactive, and self-motivated.
- A hands-on problem solver who is comfortable navigating open-ended research challenges.
- A collaborative colleague who enjoys working across disciplines.
- Comfortable working in a fast-paced, high-impact environment.
Requirements
- M.Sc., Ph.D., or equivalent industry experience.
- Strong foundation in mathematics.
- Expertise in image processing and computer vision.
- A strong research track record in generative modeling, computer vision, or a related field.
- Strong programming and software design skills.
- Hands-on experience training video diffusion models.
- Deep understanding of modern generative-model architectures and training methods.
How to stand out in the crowd
- Experience with multimodal models.
- Experience optimizing model inference.
- Trained large scale models over cloud infrastructure.
About Us
D-ID’s generative AI technology transforms customer experience, learning and development, sales, and marketing through video. Our platform enables creators to generate photorealistic digital presenters from text—dramatically reducing the cost and complexity of producing video content at scale and real time.
Our customers include most of the Fortune 1000 companies, and organizations across financial services, automotive, technology, retail, entertainment, marketing, production, and social media.
Founded in 2017 and backed by leading venture capital firms, D-ID makes its technology available through a self-service studio, API, and plug-ins. Our technology powers lifelike AI assistants and other interactive digital experiences.
To date, more than 150 million videos have been created using D-ID’s technology, and over 200,000 developers have used our API.
Now it’s your perfect time to join us!