Remote (India) · Senior · Remote
Applicants who checked fit first are 3.1× more likely to hear back
Your score for this role already exists
ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.
No credit card · 1 tap with Google
HIGH
Jobgether is actively reviewing profiles and moving candidates through the pipeline right now.
First 72 hours
Still inside it - posted 14h agoEarly applicants get seen before the pile builds.
Not a repost
The first time we've seen this listing - it hasn't been closed and reopened.
You almost certainly match several of these already. Unlock your skill map to see the matches, the gaps, and what to fix first.
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Sr Machine Learning Engineer - AI based in India.
This role focuses on developing and deploying efficient small language models (SLMs) for real-world AI applications.
You’ll work across model fine-tuning, optimization, evaluation, and production deployment.
The position combines machine learning engineering with practical MLOps and performance engineering.
You’ll help make models smaller, faster, and more efficient through techniques such as quantization, pruning, and knowledge distillation.
Your work will extend to edge devices, mobile environments, and local infrastructure where latency and resource efficiency are critical.
You’ll also build reliable pipelines and monitoring systems that support models throughout their production lifecycle.
The role offers an opportunity to contribute to advanced AI systems while solving challenging performance and deployment problems.
Fine-tune and train small language models using Hugging Face, TRL, and adapter-based techniques such as LoRA, QLoRA, and PEFT.
Optimize models for efficient inference through quantization, pruning, knowledge distillation, and other model-compression approaches.
Deploy machine learning models to edge devices, mobile platforms, and local servers while meeting demanding latency and resource constraints.
Build end-to-end MLOps pipelines covering data ingestion, experimentation, model development, evaluation, deployment, and production operations.
Establish and maintain model evaluation frameworks, benchmarking processes, and custom test suites to measure model quality and performance.
Monitor production models for accuracy, inference latency, CPU/GPU utilization, and other relevant operational metrics.
Contribute to continuous improvements in AI deployment workflows, model efficiency, reliability, and scalability.
Hands-on experience developing, training, and fine-tuning small language models or other transformer-based models using Hugging Face and related tooling.
Strong knowledge of adapter-based fine-tuning methods, including LoRA, QLoRA, and PEFT.
Practical experience with model optimization techniques such as quantization, pruning, and knowledge distillation.
Experience deploying machine learning models to edge devices, mobile environments, or local/on-premises infrastructure.
Ability to design and implement end-to-end MLOps pipelines from data ingestion through production deployment.
Experience monitoring machine learning systems in production, including model accuracy, latency, and hardware utilization.
Strong understanding of model evaluation, benchmarking, and performance optimization.
Experience with experiment tracking, model registries, and ML-focused CI/CD practices is valuable.
Knowledge of ONNX export and cross-platform inference is an advantage.
Strong problem-solving skills and the ability to work effectively in a collaborative engineering environment.
Remote working opportunity in India.
Opportunity to work on applied AI, SLMs, model optimization, and production machine learning systems.
Exposure to edge, mobile, and resource-constrained AI deployment environments.
Opportunity to work with modern machine learning and MLOps technologies.
Inclusive and collaborative workplace culture that values diverse perspectives and backgrounds.
Professional growth opportunities through work on advanced AI engineering challenges.
Free · no signup
Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.
No spam. Just jobs and resources.
Why people use ASAI
Scored, not searched. Every role ranked against your actual profile.
Alerts as often as hourly. Reach new roles while the pile is still small.
Skill gaps, spelled out. See exactly which requirements you don't meet yet.
Verified jobs, only. Say no to ghost jobs. Your time deserves respect.
Keep browsing