Jobs Companies Pika ML Engineer, Inference & Optimization

Sobre esta vaga de ML Engineer, Inference & Optimization na Pika

Pika · Presencial · Palo Alto HQ

About the Role

We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your expertise will drive significant improvements to model speed and efficiency, ensuring our creative AI systems deliver industry-leading user experiences at scale.

 

You will design and optimize inference pipelines, implement state-of-the-art acceleration techniques, and work closely with researchers and engineers across the team to push the boundaries of what’s possible in real-time AI deployment. Your efforts will play a foundational role in powering the next generation of Pika’s video and language models.

 

What You’ll Do

  • Accelerate Inference: Lead and implement advanced inference acceleration techniques, including attention optimization and quantization for efficient model serving.

  • Maximize GPU Parallelism: Engineer and optimize GPU strategies across tensor, sequence, and pipeline parallelism (TP, SP, PP) for maximal efficiency and scalability.

  • Programming for Performance: Develop and optimize high-performance computing kernels and distributed workloads using CUDA and NCCL.

  • Advance AI Deployment: Collaborate with research and engineering teams to bring state-of-the-art videogen and large language models into production.

  • Improve Training Efficiency: (Bonus) Contribute to improvements in model training speed, stability, and resource utilization as part of our deployment lifecycle.

  • Technical Excellence: Drive rigorous code reviews, participate in technical discussions, and mentor fellow engineers on best practices in inference and GPU programming.

 

What We’re Looking For

  • Experience: 5+ years engineering experience, with a strong track record in inference acceleration and model deployment at scale.

  • Inference Mastery: Proven expertise in inference optimization, including quantization, attention acceleration, and deep learning compiler stacks.

  • GPU & Parallelism: Deep knowledge of GPU programming (CUDA, NCCL) and experience with SP, TP, PP, and other forms of parallelism for distributed inference.

  • AI Domain Knowledge: Familiarity with video generation (videogen) models and large language models (LLMs).

  • Collaboration: Strong cross-discipline communication skills; able to drive shared goals across research and engineering functions.

  • Ownership Mindset: Self-driven, solutions-oriented, and capable of managing ambiguity in a fast-paced startup environment.

  • Bonus: Experience in enhancing training efficiency, stability, or resource optimization for large models.

 

Nice to Have

  • Experience with high-throughput video or real-time streaming model deployment

  • Familiarity with distributed training and optimization toolkits

  • Contributions to open source projects in AI infrastructure or deep learning compilers

  • Startup or rapid prototyping experience

 

What We Offer

  • Competitive salary in the AI industry

  • Equity in a fast-growing startup shaping the future of AI

  • Comprehensive health benefits, monthly stipends, company retreats

  • A supportive and collaborative office culture—we’re all building and launching together

 

About Pika

At Pika, we're crafting a future where video creation is seamless, intuitive, and universally accessible. Our mission is to empower creativity by breaking down technical barriers using the transformative power of AI. We’re a tight-knit, energetic team based in Palo Alto, CA, valuing efficiency, curiosity, and the ambition to make a meaningful impact on the world.

 

We work from our Palo Alto office 3–5 days a week and welcome applicants who are eager to contribute onsite.

Pronto para se candidatar à Pika?
Candidatar-se à Pika

Como este salário de ML Engineer se compara

Esta vaga paga $300,000/yracima da faixa típica para vagas de ML Engineer.

$144,286 a mediana $211,910 $300,000

Faixa típica $175,000–$255,700/yr, com base em 729 vagas de ML Engineer comparáveis na JobsRadar (pagamento anualizado em USD). Ver insights salariais de ML Engineer →

Vagas semelhantes

Rubrik Job Board
Senior Machine Learning Engineer
Rubrik Job Board
⚡ Candidate-se cedo Palo Alto, CA Presencial $188,500–$282,700
● Nova 👁 Vista ✓ Candidatada há 1d
TY
Principal ML Engineer
Typeface
⚡ Candidate-se cedo Palo Alto, CA Híbrido $230,000–$260,000
● Nova 👁 Vista ✓ Candidatada há 3m
TY
Staff ML Engineer
Typeface
⚡ Candidate-se cedo Palo Alto, CA Híbrido $190,000–$234,000
● Nova 👁 Vista ✓ Candidatada há 3m
Stitch Fix
ML Platform Engineer
Stitch Fix
⚡ Candidate-se cedo Remote, USA · local restrito $136,000–$167,000
● Nova 👁 Vista ✓ Candidatada há 2h
MetroStar
Sr. AI/ML Engineer III (6707)
MetroStar
⚡ Candidate-se cedo Tysons Corner, VA Presencial $212,000–$248,000
● Nova 👁 Vista ✓ Candidatada há 2h
MetroStar
Sr. AI/ML Engineer III (6707)
MetroStar
⚡ Candidate-se cedo Herndon, VA Presencial $212,000–$248,000
● Nova 👁 Vista ✓ Candidatada há 2h
Roku
Senior Machine Learning Engineer
Roku
⚡ Candidate-se cedo Austin, Texas Presencial
● Nova 👁 Vista ✓ Candidatada há 4h
Roku
Senior Machine Learning Engineer
Roku
⚡ Candidate-se cedo San Jose, California Presencial $229,500–$367,100
● Nova 👁 Vista ✓ Candidatada há 4h
Roku
Senior Machine Learning Engineer, Content Platform
Roku
⚡ Candidate-se cedo Bengaluru, India Presencial
● Nova 👁 Vista ✓ Candidatada há 4h

Cadastre-se para receber sugestões sob medida com base nas vagas que você abre e nas buscas que você salva.

Mais vagas na Pika

Ver todas as vagas na Pika →

Candidatar-se agora
🤖

Opa — calma aí

A JobsRadar foi feita para pessoas de verdade passando por um momento difícil na busca por emprego — não para requisições automatizadas. Você está clicando rápido demais e agora está temporariamente bloqueado.

Volte mais tarde. Se você está mesmo procurando emprego, estamos com você — apenas aja como um ser humano.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Ganhe vantagem na sua busca por emprego.

Entre no nosso canal do Telegram para o que ajuda você a conseguir a vaga — referências salariais, o pulso semanal do mercado e avisos de novos recursos. Sem spam, só sinal.

Entre no canal — é grátis