Jobs Companies Middesk Lead Data Scientist

Sobre este puesto de Lead Data Scientist en Middesk

Middesk · Híbrido · San Francisco

About Middesk:

Middesk is building the data and intelligence infrastructure that helps businesses work together with confidence. We started by creating a comprehensive platform for understanding businesses, bringing together authoritative and proprietary data to help customers verify business identities, onboard customers faster, and manage risk throughout the customer lifecycle.

Today, Middesk is used by more than 700 banks and fintechs, and in 2025 we verified more than 7 million companies. We've also expanded beyond business verification to help companies form, register, manage, and maintain their businesses, supporting more than 50,000 companies in setting up over 100,000 accounts required to hire employees, run payroll, and stay compliant.

Middesk came out of Y Combinator, and is backed by Sequoia Capital, Accel, Insight Partners, and Canapi. We're proud to be named on the Forbes Fintech 50 and Best Startup Employers lists.

About The Role:

We are actively building AI-driven applications that streamline customer workflows, focusing on business onboarding. With our proprietary identity data assets and deep domain expertise, we are uniquely positioned to expand into a broader set of AI-powered solutions that drive long-term growth.

We’re looking for a hands-on applied ML expert to help build the technical foundation for these efforts. Ideally you have shipped external-facing models in the risk/fraud space and know the messy realities of imbalanced data, low labels, and changing behavior. This is a highly technical, hands-on role with wide influence on how we design, build, and scale ML at Middesk.

We follow a hybrid work model, and for this role, there is an expectation of 2 days per week in our SF/NYC office. Candidates should be based within a commutable distance, as we believe in the value of in-person collaboration and building strong team connections while also supporting flexibility where possible.

What You'll Do:

  • Build risk & fraud ML applications: Deliver production ML models in fraud, trust & safety, KYB, and compliance domains, with measurable impact on customer workflows.

  • Tackle hard data problems: Work on classification problems with extreme class imbalance, sparse signals, and “cold start” label challenges.

  • Innovate in feature engineering & labeling: Use graph-based techniques, weak supervision, LLMs, and AI agents to improve signal extraction and automate labeling process.

  • Establish ML infrastructure foundations: Partner with the ML infra team to design feature services, model training pipeline, model serving standards, and orchestration to scale multiple ML use cases.

  • Design and implement knowledge graph solutions: Leveraging LLMs for graph construction, querying, and retrieval to enhance entity resolution and business identity use cases.

What We're Looking For:

  • 7+ years of production ML experience in one or more of the following areas:

    • Building Production ML for risk, fraud, credit, or trust & safety: Track record of shipping external-facing ML applications in one or more of these domains.

    • Knowledge graph applications: Hands-on experience building, querying, or extracting signals from knowledge graphs—ideally over business entity networks (companies, persons, addresses, relationships) to support identity verification, fraud detection, or risk decisioning.

    • Entity resolution for business or individual identities: Experience disambiguating and linking records across noisy, incomplete, or conflicting data sources—particularly in KYB, KYC, AML, or identity verification contexts where the same real-world entity may appear under different names, addresses, or tax IDs.

  • Expertise in classification with real-world ML challenges, for example: imbalanced labels, sparse signals, cold start, and production version management.

  • Hands-on ML infrastructure experience: feature stores, model management, ML training/serving pipelines.

  • Comfort as a senior IC: setting technical direction, mentoring peers, and establishing best practices.

Nice-To Have:

  • B2B SaaS experience, ideally building ML products for enterprise customers.

  • ML pipeline and automation engineering: Experience building end-to-end training harnesses that automate feature engineering, data validation, and model training.

  • Experience scaling ML across multiple products or risk domains.

¿Listo para postularte en Middesk?
Postúlate en Middesk

Cómo se compara este salario de Data Scientist

Este puesto paga $230,000/yren línea con el rango típico para los puestos de Data Scientist.

$165,000 la mediana de $225,000 $325,080

Rango típico $197,500–$264,500/yr, a partir de 180 ofertas comparables de Data Scientist en JobsRadar (salario anualizado en USD). Ver datos salariales de Data Scientist →

Empleos similares

Zoomlogi
Founding Data Scientist
Zoomlogi
⚡ Postúlate pronto San Francisco HQ Híbrido $180,000–$260,000
● Nuevo 👁 Visto ✓ Postulado hace 5h
Triumph
Data Scientist
Triumph
⚡ Postúlate pronto San Francisco HQ Presencial $200,000–$400,000
● Nuevo 👁 Visto ✓ Postulado hace 5h
Notion
Data Scientist, Growth
Notion
⚡ Postúlate pronto San Francisco, California Híbrido $145,000–$233,000
● Nuevo 👁 Visto ✓ Postulado hace 5h
Gridware
Senior Applied Scientist - Multi-Sensor & On-Device ML
Gridware
⚡ Postúlate pronto San Francisco, CA Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 6h
Reddit
Staff Data Scientist, Marketing
Reddit
⚡ Postúlate pronto Remote - United States · restringido por ubicación $217,000–$303,900
● Nuevo 👁 Visto ✓ Postulado hace 8h
Reddit
Staff Data Scientist, Marketing
Reddit
⚡ Postúlate pronto Remote - Ontario, Canada · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 8h
Reddit
Staff Data Scientist, Consumer
Reddit
⚡ Postúlate pronto Remote - Ontario, Canada · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 8h
Lyft
Staff Applied Scientist
Lyft
⚡ Postúlate pronto San Francisco, CA Híbrido $193,600–$242,000
● Nuevo 👁 Visto ✓ Postulado hace 8h
Eight Sleep
Data Scientist, Growth
Eight Sleep
⚡ Postúlate pronto Remote Global · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 13h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Middesk

Ver todos los empleos en Middesk →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis