← Back to jobs
BBlue Machines AI

ML Researcher / Engineer - Voice and Speech to Speech Models

Blue Machines AI

Bengaluru
Full-Time
0-3 Years experience

Description

Location: Bengaluru (Work from Office - Domlur)

Team: AI & Machine Learning

Experience: 2–7 years

What You'll do:

  • Fine-tune and deploy LLMs, TTS, STT, and voice models for use in real-time conversations with millions of users.

  • Convert unstructured, messy real-world audio/text data into clean, high-quality datasets for training and evaluation.

  • Build inference pipelines optimized for low-latency, high-accuracy voice agents and multimodal interfaces.

  • Work closely with infra and product teams to ship production-grade GenAI models with observability, fallback, and monitoring.

  • Experiment with GANs, diffusion models, audio generation, and multimodal fusion to power next-gen AI agents.

  • Own the full model lifecycle — from research and training to deployment, testing, and iteration.

What we're Looking for:

  • 2-7 years of hands-on experience in AI / ML roles, ideally in startups or product-driven teams.
  • Strong grasp of LLM fine-tuning, instruction tuning, or pretraining techniques.
  • Familiarity with TTS/STT systems, Whisper, Tacotron, VITS, or other open source models .
  • Experience with multimodal architectures, generative audio, GANs, or diffusion-based models.
  • Ability to work with real-world messy data, design training pipelines, and debug model failure modes.
  • Fluency in frameworks like PyTorch, HuggingFace, TensorFlow, and ecosystem tools (ONNX, Triton, LangChain, etc.).
  • Passion for building high-impact AI features that ship to real customers.

About Blue Machines AI

-

Industry: no-mentionEmployees: 10+Website