← Back to results

Research Scientist (diffusion)

Join Genmo as a Research Scientist to innovate in diffusion models for text-to-video generation.

Location
San Francisco HQ
Compensation
Not disclosed
Level
senior
Type
full time · On-site

Posted by employer 3 days ago

First seen on Joblaze 1 week ago

Last verified on the company career page 1 day ago

Apply at Genmo → Save job Scanned from genmo.ai

Requirements

Experience
3+ years
Education
PhD

Not disclosed in this posting: compensation, visa sponsorship.

Joblaze summary

In this role, the Research Scientist focuses on advancing diffusion models to convert text into high-quality video content, driving innovation through the development of new algorithms and architectures. Candidates should possess a Ph.D. in a relevant field, with a strong publication record in generative models and proficiency in Python and deep learning frameworks like PyTorch or TensorFlow. This position is ideal for experienced researchers with a background in generative AI, particularly those who have worked on text-to-video projects and have a collaborative mindset. Genmo emphasizes a culture of innovation and encourages contributions to the research community.

Joblaze insights

Quick facts

Is the Research Scientist (diffusion) role remote?
No — this is an on-site role in San Francisco HQ.
How much experience is required?
At least 3 years of relevant experience for this Research Scientist (diffusion) role.
Where is the role based?
Genmo is hiring for this position in San Francisco HQ.
What's the tech stack?
Joblaze extracted these technologies from the posting: Diffusion Models, PyTorch, Python, TensorFlow, generative models, text-to-video generation.
What seniority level is this role?
Genmo targets senior candidates for this position.
Is this full-time or contract?
Full-time for this Research Scientist (diffusion) role at Genmo.

From the original posting

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

Role overview:

We are seeking an exceptional Research Scientist to join our team, focusing on developing cutting-edge diffusion models for text-to-video generation. In this role, you will be at the forefront of innovation, creating novel architectures and algorithms that transform written descriptions into stunning, coherent video content.

Key responsibilities:

  • Lead research initiatives in advanced diffusion models for text-to-video generation, focusing on improving visual quality, temporal consistency, and semantic fidelity

  • Develop and implement state-of-the-art algorithms for translating textual descriptions into dynamic video content

  • Design and conduct rigorous experiments to validate new ideas and evaluate model performance

  • Collaborate with cross-functional teams to integrate research breakthroughs into our production pipeline

  • Stay at the cutting edge of the field by regularly reviewing academic literature and attending top-tier conferences

  • Contribute to the research community through high-quality publications and open-source contributions

  • Mentor junior researchers and foster a culture of innovation within the research team

  • Work closely with product teams to align research directions with user needs and market opportunities

Qualifications:

  • Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a closely related field

  • Must have:

    • Strong publication record in top-tier conferences (e.g., CVPR, ICCV, NeurIPS, ICML) with a focus on generative models, particularly diffusion models

    • Extensive experience implementing and optimizing large-scale generative models for image or video tasks

    • Deep understanding of state-of-the-art techniques in text-to-image and text-to-video generation

    • Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow

    • Excellent communication skills with the ability to explain complex technical concepts to diverse audiences

    • Proven ability to work collaboratively in a team environment

  • Ideal candidate will have:

    • Postdoctoral or industrial research experience in generative AI for video

    • Hands-on experience with text-to-video generation projects

    • Expertise in other generative model architectures (e.g., GANs, VAEs) and their applications to video

    • Experience working with large-scale datasets and distributed computing environments

    • Track record of successful collaboration with product teams on technology transfers

    • Familiarity with video codecs, compression techniques, and perceptual quality metrics

    • Contributions to open-source projects in the field of generative AI

Additional information

The role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.

Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.

Similar positions

Genmo
Research Engineer (New Grad)
Genmo · San Francisco HQ
Genmo
Research Scientist (post-training)
Genmo · San Francisco HQ
Genmo
Founding Product Designer
Genmo · San Francisco HQ
World Labs
Research Scientist (Generative Modeling)
World Labs · San Francisco