Skip to main content
Weave

Staff ML - GenAI Engineer, Voice & Speech

RemoteIndia only
Published
Role
AI / ML
Experience
Staff
Employment
Full-time
Salary not disclosed
Check eligibility

Open to IN only. Set where you work from to check your eligibility.

No BS summary

Staff-level ML/AI engineer with 15+ years of experience, focused on voice/audio GenAI at scale. Must have deep LLM/RAG/prompt engineering/fine-tuning, audio/voice models, production ML, Python/ML tooling, distributed systems, cloud, and large-scale data experience. Fully remote but must be located in India and overlap India and US business hours.

Core skills

LLMsRAGAudio/voice models

Required skills

Prompt EngineeringFine-tuningLLM evaluationsPythonJupyterDagsterMLFlowKubeFlowDVCTriton ServerPostgreSQLmulti-modal modelsAWSGCP

Optional skills

Model Context ProtocolKubernetesGKEOperator PatternGitOpsIaC

What you'll do

  • Design and develop machine learning infrastructure, tooling, and models to help teams deliver world-class experiences.
  • Help product and development teams understand the data lifecycle and the inherent experimental nature of machine learning.
  • Build internal products and platforms to enable teams to incorporate AI into their features and customer-facing products.
  • Consult with teams to help them understand common patterns, anti-patterns, and tradeoffs of machine learning.
  • Guide teams through creating excellent customer experiences end to end.
  • Build scalable, resilient services to support data integration, event processing, and platform extensions.
  • Contribute to the continued evolution of product functionality that services large amounts of data and traffic.
  • Write high-quality, performant, sustainable, and testable code.
  • Hold yourself accountable for the quality of the code you produce.
  • Coach and collaborate inside and outside the team.
  • Help others grow by sharing expertise and encouraging best practices.
  • Work in a cloud environment, considering implementation through distributed components and services.
  • Work with stakeholders to translate product goals into actionable engineering plans.
  • Build models for new products with emerging technologies at scale.

What they require

  • High integrity, team-focused approach, and collaboration skills to build tight-knit relationships across Weave with various roles and stakeholders.
  • Responsive person with a strong bias for action.
  • 15+ years of experience in Machine Learning or AI, with a focus on or expertise in audio and voice GenAI solutions at scale.
  • Deep expertise with modern ML tools and techniques such as LLMs, RAG, Prompt Engineering, Fine-tuning, high-scale audio/voice models, and LLM evaluations.
  • Experience moving and storing TBs of data or 100Ms to 110B records.
  • Experience building and deploying ML-driven B2B multi-tenant applications in production environments at scale for external products and customers.
  • Experience with common ML technologies such as Python, Jupyter, Workflow Engines, DVC, Triton Server, LLMs, Postgres, and others.
  • Experience with modern ML tools and techniques such as LLMs, RAG, Prompt Engineering, Fine Tuning, LLM evaluations, multi-modal models, and others.
  • Experience with data labelling or annotation for audio or text use cases.
  • Understanding of distributed systems and building scalable, redundant, and observable services.
  • Expertise in designing systems for distributed data sets and services.
  • Experience building solutions to run on one or more public clouds.
  • Experience providing stable, well-designed libraries and SDKs for internal use.
  • Self-driven and a thirst for learning in a quickly changing industry.
  • Demonstrated track record of delivering complex projects on time and experience working in enterprise-grade production environments.
  • Demonstrated capacity for leadership or mentorship.
  • Strategic thinker with a strong technical aptitude and a passion for execution.
  • Employment is contingent upon successful completion of a background check.
  • Must overlap with India and US business hours.
  • Preferred: A background with data analysis, visualisation, and presentation.
  • Preferred: 14+ years of experience in engineering and systems with strong proficiency in coding and system design.
  • Preferred: Experience with low-latency natural language models and pipelines at scale.
  • Preferred: Experience with real-time audio models and voice use cases such as transcription, ASR pipelines with interruption detection, audio alignment, and speech synthesis.
  • Preferred: Experience with emerging technologies such as Model Context Protocol (MCP).
  • Preferred: Proficient understanding of containers, orchestrators, and usage patterns at scale.
  • Preferred: Experience with Kubernetes or GKE and the Operator Pattern (GCP).
  • Preferred: Experience with highly sensitive data such as PHI (HIPAA) and PII data.
  • Preferred: Experience with automation and container-based workflow engines.
  • Preferred: Experience with GitOps, IaC, and configuration-driven systems.
  • Preferred: A preference for open source solutions.
  • Preferred: A track record of clean abstractions and simple-to-use APIs.
  • Preferred: A desire to advance the state of the art with new and innovative technologies.
  • Preferred: Enjoys working in a greenfield environment using rapid prototyping.
  • Preferred: Enjoys working with open-ended, evolving problems and domains.

Benefits

  • Fully remote opportunity in India.

Weave provides an all-in-one platform for healthcare practices, supporting scheduling, payments, communication, and reviews to improve patient experience and practice efficiency.

🇺🇸 United StatesHealthcareEnterprise

Details

Apply routeDom
Salary not disclosed