Group Product Manager - Inference and Token Factory

Added
3 hours ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

model serving latency inference serving gpu compute cost efficiency

πŸ“‹ Description

  • Own product strategy and roadmap for inference infrastructure and token-serving
  • Lead and develop PMs on inference and serving infrastructure
  • Partner with platform engineering to boost throughput, latency, and cost efficiency
  • Track trends in inference workloads and translate to roadmap priorities
  • Define and monitor metrics for inference performance and efficiency
  • Represent inference and token-serving priorities in cross-functional planning
  • Work with key customers to understand evolving inference needs

🎯 Requirements

  • PM experience in AI/ML infrastructure or inference serving essential
  • Experience managing PMs or leading a product team
  • Strong technical understanding of model serving, inference optimization, or GPU compute
  • Track record shipping infrastructure products in fast-evolving space
  • Excellent prioritization and cross-functional communication
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Product Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Product Jobs

See more Product jobs β†’