Staff Research Engineer - Multimodal Generative Modelling

Added
3 hours ago
Location
Type
Full time
Salary
Salary not provided

Related skills

pytorch streaming llm diffusion transformer

📋 Description

  • Shape roadmap to unlock new model capabilities for customers.
  • Propose novel multi-modal architectures (text and voice).
  • Develop streaming, conversational systems for low-latency voice-video synthesis.
  • Design solutions that reinforce expressive, natural interaction.
  • Implement designs from pretraining through post-training.
  • Integrate architectures like neural codecs and diffusion to boost realism.

🎯 Requirements

  • Strong understanding of generative modelling for multimodal data.
  • Hands-on experience with LLMs or transformer architectures.
  • Proficiency in PyTorch, including distributed training and optimization.
  • Solid grasp of time-series modeling and tokenization for audio/speech/video.
  • Proven ability to prototype quickly, test hypotheses, and iterate.
  • Experience shipping a generative model into live product at meaningful scale.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →