Join D-ID's research team to develop the next generation of multimodal video diffusion models and bring cutting-edge research into real-time products used at global scale.
About The Team
Our research team is small, fast-moving, and highly impactful. We develop multimodal generative models for real-time video creation, turning cutting-edge research into scalable products.
We are looking for researchers with deep expertise in generative video to join a team at the core of D-ID's technology and help shape the future of generative AI.
Why Work at D-ID?
Work on cutting-edge video-generation technology at the forefront of AI
See your research become part of products used at global scale
Own meaningful research and engineering challenges as part of a small, high-impact team
Help solve complex problems in real-time video generation
Build technology that brings the next generation of AI assistants to life
Role & Responsibilities
Research, develop, and evaluate generative AI models for video generation
Design experiments and evaluation methods covering visual quality, temporal consistency, identity preservation, and performance
Explore, validate, and implement innovative research ideas
Own the full lifecycle from research and experimentation to production deployment
Build models and systems that perform reliably at scale and in real time
Collaborate closely with research, engineering, and product teams
You are
A curious researcher and dedicated self-learner
Independent, proactive, and self-motivated
A hands-on problem solver who is comfortable navigating open-ended research challenges
Comfortable working in a fast-paced, high-impact environment
Requirements
M.Sc., Ph.D., or equivalent industry experience
Strong foundation in mathematics
Expertise in image processing and computer vision
A strong research track record in generative modeling, computer vision, or a related field
Strong programming and software design skills
Hands-on experience training video diffusion models
Deep understanding of modern generative-model architectures and training methods
How to stand out in the crowd
Experience with multimodal models
Experience optimizing model inference
Trained large scale models over cloud infrastructure.
About Us
D-ID's generative AI technology transforms customer experience, learning and development, sales, and marketing through video. Our platform enables creators to generate photorealistic digital presenters from text—dramatically reducing the cost and complexity of producing video content at scale and real time.
Our customers include most of the Fortune 1000 companies, and organizations across financial services, automotive, technology, retail, entertainment, marketing, production, and social media.
Founded in 2017 and backed by leading venture capital firms, D-ID makes its technology available through a self-service studio, API, and plug-ins. Our technology powers lifelike AI assistants and other interactive digital experiences.
To date, more than 150 million videos have been created using D-ID's technology, and over 200,000 developers have used our API.
Now it's your perfect time to join us!