Staff ML Engineer, GenAI, Voice & Speech

Posted 6hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Staff ML Engineer building scalable voice and speech GenAI infrastructure for Weave’s AI-powered products. Designing models, platforms, and distributed services for high-volume customer experiences.

Responsibilities:

  • Design and develop machine learning infrastructure, tooling, and models for product teams
  • Help teams understand the machine learning data lifecycle and experimental process
  • Build internal products and platforms enabling AI features in customer-facing products
  • Consult teams on machine learning patterns, anti-patterns, tradeoffs, and end-to-end customer experiences
  • Build scalable, resilient services for data integration, event processing, and platform extensions
  • Evolve product functionality serving large volumes of data and traffic
  • Write high-quality, performant, sustainable, and testable code
  • Coach and collaborate with stakeholders inside and outside the team
  • Design functionality across distributed cloud components and services
  • Translate product goals into actionable engineering plans

Requirements:

  • 15+ years of experience in Machine Learning or AI, focused on or with expertise in audio and voice GenAI solutions at scale
  • Deep expertise with LLMs, RAG, prompt engineering, fine-tuning, high-scale audio/voice models, and LLM evaluations
  • Experience moving and storing terabytes of data or hundreds of millions to billions of records
  • Experience building and deploying production ML-driven B2B multi-tenant applications at scale
  • Experience with Python, Jupyter, workflow engines such as Dagster, MLFlow, and KubeFlow, DVC, Triton Server, LLMs, and Postgres
  • Experience with data labelling or annotation for audio or text use cases
  • Understanding of distributed systems and scalable, redundant, observable services
  • Expertise designing systems for distributed datasets and services
  • Experience deploying solutions on public clouds such as AWS or GCP
  • Experience providing stable libraries and SDKs for internal use
  • Demonstrated delivery of complex projects in enterprise-grade production environments
  • Demonstrated leadership or mentorship capacity
  • Experience with low-latency natural language models and pipelines at scale
  • Experience with real-time audio models and voice use cases, including transcription, ASR pipelines with interruption detection, audio alignment, and speech synthesis
  • Experience with Model Context Protocol (MCP)
  • Understanding of containers and orchestrators at scale; Kubernetes or GKE and the Operator Pattern are a plus
  • Experience with sensitive PHI/HIPAA and PII data
  • Experience with automation and container-based workflow engines
  • Experience with GitOps, IaC, and configuration-driven systems
  • Strong coding and system design proficiency
  • High integrity, collaboration, responsiveness, bias for action, self-direction, strategic thinking, and technical aptitude
  • Successful completion of a background check

Benefits:

  • Fully remote opportunity in India
  • Opportunity to work on AI-powered applications and emerging technologies
  • Coaching and collaboration opportunities
  • Inclusive workplace and equal opportunity employer
  • Accommodation support for disabilities or special needs