
Process Nine Technologies
Machine Learning Lead
GurugramPosted Today₹2,500,000 – ₹3,000,000
Full TimeSeniorIN
See how this job matches your profile
Sign in for an AI-powered fit score, breakdown, and a tailored resume.
Job Description
ML Leads JDKey ResponsibilitiesModel Training & Fine-Tuning: Build, fine-tune, and optimize state-of-the-art NLP, LLM, Speech, and Vision models for scheduled Indian languages, utilizing parameter-eff
Key Highlights
- Model Training & Fine-Tuning: Build, fine-tune, and optimize state-of-the-art NLP, LLM, Speech, and Vision models for scheduled Indian languages, utilizing parameter-efficient methods (LoRA, QLoRA, PEFT).
- Indic Tokenization & Linguistics: Architect custom tokenizers and text-normalization pipelines to address the "fertility problem" in Devanagari, Dravidian, and other regional scripts, ensuring low-latency and cost-effective model inference.
- Multimodal System Design: Develop robust OCR engines capable of parsing complex script geometries (conjoint consonants, Shirorekha, vowel modifiers) and integrate them into document intelligence pipelines.
- Speech Engineering: Deploy and scale robust STT (Speech-to-Text) and TTS (Text-to-Speech) pipelines capable of handling heavy code-mixing (e.g., Hinglish, Tanglish), regional accents, and localized dialects.
- Vernacular Guardrails & Evaluation: Establish culturally contextual benchmark datasets and implement safety guardrails.
Qualifications
Required Qualifications
- Education: Bachelor’s or Master's degree in Computer Science, Mathematics, Statistics, or a closely related quantitative field.
- Experience: 4+ years of professional experience building and deploying machine learning models in production environments, with a proven track record in Indian Language NLP, Speech, or Anomaly Detection.
- Programming: Expert-level proficiency in Python and standard ML frameworks (PyTorch, TensorFlow).
- Indic AI Stack: Direct, hands-on experience with specialized Indic frameworks and datasets (e.g., AI4Bharat's IndicTrans2/IndicWhisper, Bhashini API, Kathbath, Sarvam-105B, or Aksharantar).
- Fraud Stack: Proficiency in tabular/graph-based ML toolkits (XGBoost, LightGBM, PyTorch Geometric) and handling highly imbalanced target variables (SMOTE, class weights).
- NLP & LLMs: Deep understanding of Transformer architectures, sequence-to-sequence modeling, cross-lingual embeddings, vector databases (Milvus, Pinecone, Qdrant), and quantization tools (bitsandbytes, GPTQ).
- Speech & Vision Processing: Experience processing raw audio signals (grapheme-to-phoneme conversion, spectrogram analysis) or document structures using OCR networks (CRAFT, DBNet, LayoutLM).
- Handling Code-Mixing: Proven ability to build models that gracefully parse text or speech containing heavy code-switching (mixed Latin/regional scripts, multi-language grammar).
Skills & Technologies
Machine Learning (ML)Natural Language Processing (NLP)TensorFlowDeep LearningPyTorchFine-tuning LLMsLoRA / QLoRASpeech-to-Text (STT)Text-to-Speech (TTS)OCRMLOps
About the Company
Process Nine Technologies
View company profile →
Interested in this role?
Sign in or create a free account to see how this job matches your skills, apply with one click, and let our AI tailor your resume.
Sign in to applyAI-powered resume optimization
Save and track your applications
Job Details
Employment Type
Full Time
Experience Level
Senior
Salary Range
₹2,500,000 – ₹3,000,000
Location
Gurugram
Posted
Today
Country
IN