Jobs Companies Plaud Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco

Sobre esta vaga de Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco na Plaud

Plaud · Híbrido · San Francisco, CA

About Plaud Inc.

Plaud is building the real-world AI interface for professionals to amplify intelligence, elevate productivity and performance, loved by over 2,000,000 users worldwide since 2023. With a mission to amplify human intelligence, Plaud captures, structures, and compounds the intelligence generated in conversations — so humans can think better, decide faster, and execute with clarity.

 

Plaud Inc. is a Delaware-incorporated, San Francisco-based company pushing the boundary of human–AI intelligence through a hardware–software combination. With full ISO 27001, ISO 27701, SOC 2, GDPR, EN18031, and HIPAA compliances, Plaud is committed to the highest standards of data security and privacy protection.

To learn more about Plaud, please visit https://www.Plaud.ai and follow along on Instagram, X, Facebook, LinkedIn, and YouTube

 

Why You Should Join Us

Plaud is building the next generation intelligence infrastructure and interfaces to capture, extract, and utilize intelligence from what people say, hear, see, and think.

  • Plaud is a bootstrapped, skyrocketing, profitable company with a $250M revenue run rate achieved in just three years.

  • Define the next-gen paradigm for human-AI interaction.

  • Gain exposure to cutting-edge AI for Pro tools and play a direct role in our global expansion.

  • Work with passionate teammates who value innovation, collaboration, and customer success.

  • Grow your career in a culture that champions continuous learning and fast career development.

  • Market-competitive compensation, global exposure, and a vibrant, creativity-fueled work atmosphere.

 

You may be a good fit if you:

  • Have a passion for turning ambiguous, subjective concepts like a voice's naturalness, expressiveness, or conversational cadence into clear, defensible, and automated metrics that researchers and leadership can rely on.

  • Possess strong software engineering skills (especially in Python) and have experience building reliable distributed systems, data pipelines, or evaluation harnesses that can run at scale against live model checkpoints.

  • Can deeply partner with ML researchers to define exactly what "good" looks like for a Speech LLM, translating capabilities (like ASR robustness in noisy environments or TTS emotional steerability) into measurable benchmarks.

  • Are comfortable building and owning dashboards that track model health during training, improving signal-to-noise ratios, reducing evaluation latency, and making performance regressions impossible to miss.

  • Rapidly debug anomalous mid-training results to determine if a drop in performance stems from the model architecture, corrupted data, or infrastructure.

  • Communicate complex statistical results and model behaviors clearly to both technical and non-technical stakeholders.

 

Strong candidates may also have experience with:

  • Speech Metrics: Deep familiarity with both traditional (WER, CER, PESQ, etc) and modern audio evaluation frameworks (automated MOS scoring).

  • LLM-as-a-Judge: Using frontier models or finetune multi-modal LLMs to evaluate the conversational logic, transcription accuracy, audio quality, and reasoning of audio models.

  • Human Evaluation: Managing large-scale crowdsourcing operations or preference data collection to support RLHF/DPO efforts.

  • Observability: A strong background in statistics and experimental design, paired with experience building trusted tracking dashboards (e.g., Weights & Biases, MLflow).

  • Adversarial Datasets: Curating complex datasets to test edge cases, such as heavy accents, overlapping speech, or highly noisy acoustic environments.

 

What We Offer

  • Founding Team Initiative: Opportunity to be an early, foundational member of our core SpeechLLM lab, with meaningful ownership and impact on a fast-growing startup.

  • Competitive Compensation: $200K - $365K base salary + performance bonus + Equity.

  • Comprehensive Benefits: Top-tier healthcare for employees and dependents, including dental and vision, and a generous employer subsidy.

  • Retirement Planning: 401(k) plan for full-time employees with company matching.

  • Paid Time Off: Unlimited PTO, plus 13 paid holidays.

  • New Parent Leave: 12 weeks of paid time off to spend time with your new family, regardless of gender.

  • Hybrid Office: Minimum of 3x in-office per week to foster highly collaborative, fast-paced research.

  • Gear & Perks: Choice of top-of-the-line laptops/workstations, annual offsites, and a fully stocked office.

 

Plaud is and will continue to be an equal opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristics.

Pronto para se candidatar à Plaud?
Candidatar-se à Plaud

Como este salário de ML Engineer se compara

Esta vaga paga $282,500/yrem linha com da faixa típica para vagas de ML Engineer.

$170,000 a mediana $252,000 $350,250

Faixa típica $204,500–$296,200/yr, com base em 93 vagas de ML Engineer comparáveis na JobsRadar (pagamento anualizado em USD). Ver insights salariais de ML Engineer →

Vagas semelhantes

Stitch Fix
ML Platform Engineer
Stitch Fix
⚡ Candidate-se cedo Remote, USA · local restrito $136,000–$167,000
● Nova 👁 Vista ✓ Candidatada há 6h
Redwood Materials
Software Engineer - ML/Computer Vision (Battery Sorting)
Redwood Materials
⚡ Candidate-se cedo McCarran, NV; San Francisco, C... Presencial $152,500–$287,500
● Nova 👁 Vista ✓ Candidatada há 8h
Discord
Senior Software Engineer, Machine Learning (Ads)
Discord
⚡ Candidate-se cedo San Francisco Bay Area Presencial
● Nova 👁 Vista ✓ Candidatada há 11h
Figma
Software Engineer, Machine Learning
Figma
⚡ Candidate-se cedo San Francisco, CA • New York,... Presencial $153,000–$376,000
● Nova 👁 Vista ✓ Candidatada há 15h
Liftoff
Software Engineer, ML Data
Liftoff
⚡ Candidate-se cedo San Francisco Bay Area Presencial $180,000–$230,000
● Nova 👁 Vista ✓ Candidatada há 1d
Anthropic
Machine Learning Infrastructure Engineer, Safeguards Research
Anthropic
⚡ Candidate-se cedo San Francisco, CA | New York C... Presencial $350,000–$500,000
● Nova 👁 Vista ✓ Candidatada há 1d
Lyft
ML Software Engineer, ETA
Lyft
⚡ Candidate-se cedo San Francisco, CA Híbrido $140,800–$176,000
● Nova 👁 Vista ✓ Candidatada há 2d
Faire
Senior Staff Machine Learning Platform Engineer
Faire
⚡ Candidate-se cedo Kitchener-Waterloo, ON; Toront... Presencial $248,000–$341,000
● Nova 👁 Vista ✓ Candidatada há 2d
Faire
Staff Machine Learning Platform Engineer
Faire
⚡ Candidate-se cedo San Francisco, CA Presencial $246,500–$339,000
● Nova 👁 Vista ✓ Candidatada há 2d

Cadastre-se para receber sugestões sob medida com base nas vagas que você abre e nas buscas que você salva.

Mais vagas na Plaud

Ver todas as vagas na Plaud →

Candidatar-se agora
🤖

Opa — calma aí

A JobsRadar foi feita para pessoas de verdade passando por um momento difícil na busca por emprego — não para requisições automatizadas. Você está clicando rápido demais e agora está temporariamente bloqueado.

Volte mais tarde. Se você está mesmo procurando emprego, estamos com você — apenas aja como um ser humano.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Ganhe vantagem na sua busca por emprego.

Entre no nosso canal do Telegram para o que ajuda você a conseguir a vaga — referências salariais, o pulso semanal do mercado e avisos de novos recursos. Sem spam, só sinal.

Entre no canal — é grátis