ML Research Engineer - Post-training (LLMs)
ekacare· Data Science.
About this role
ML Research Engineer; Post-training (LLMs)
Bengaluru · Full-time · Experience: 4–8 yrs
India's healthcare runs in twenty-two languages, on handwritten prescriptions and ten-minute consults, and the models that should serve it are trained on the English internet. We're fixing that, in the open.
About EkaCare
EkaCare is India's connected healthcare platform: an EMR that doctors run their practices on, a personal health record used by millions of Indians, and one of the deepest integrations with India's ABDM digital-health rails. Our Parrotlet family of medical models already serves Indian doctors in production, and we open-source our work where it counts.
The role
Post-training is where a base model becomes a doctor's tool, and where most medical models quietly fail. You'll turn a strong pre-trained base into a model that follows instructions, knows what it doesn't know, stays safe in a clinical setting, and does it in a dozen Indian languages.
What you'll do
- Design SFT/IFT data mixes and chat templates for clinical tasks, and reformat the world's medical data to match.
- Run preference optimisation (DPO/ORPO/GRPO-class) and reward-model training; own the ablation grid.
- Build RL loops with verifiable medical rewards, with practising physicians in the loop. Real doctor-in-the-loop, not proxy labels.
- Create agentic and tool-use training data for healthcare workflows.
- Live in the eval → error-analysis → iterate loop; red-team your own model before the world does.
What we look for
- 2–4 years in ML with hands-on post-training of ≥7B open-weights models; you've shipped SFT plus at least one preference-optimisation method end to end, not a notebook demo.
- Fluency with the open post-training stack (TRL / NeMo RL / OpenRLHF; vLLM for rollouts).
- Strong empirical taste: ablation discipline, LLM-judge literacy, contamination paranoia.
- Solid engineering, Python, distributed-training basics, comfort in a fast codebase.
Bonus
- PPO/GRPO at scale; reward-hacking war stories.
- Multilingual or medical/clinical alignment work.
- Open-source contributions people actually use.
Frequently Asked Questions
Is the salary disclosed for the ML Research Engineer - Post-training (LLMs) position at ekacare?
The salary for this ML Research Engineer - Post-training (LLMs) role at ekacare is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the ML Research Engineer - Post-training (LLMs) position at ekacare located?
This ML Research Engineer - Post-training (LLMs) role at ekacare is based in Bengaluru, KA, India. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the ML Research Engineer - Post-training (LLMs) role at ekacare full-time or part-time?
This is listed as a FULL TIME position. It is posted as a ML Research Engineer - Post-training (LLMs) role in the Data Science. department at ekacare.
Which team or department does the ML Research Engineer - Post-training (LLMs) at ekacare belong to?
This ML Research Engineer - Post-training (LLMs) position is part of the Data Science. department at ekacare. See the full job description for more information about the team structure and responsibilities.
How do I apply for the ML Research Engineer - Post-training (LLMs) position at ekacare?
Click the "Apply Now" button on this page. You will be redirected to ekacare's official application portal hosted on keka where you can submit your application directly.
When was the ML Research Engineer - Post-training (LLMs) job at ekacare posted?
This ML Research Engineer - Post-training (LLMs) position at ekacare was posted on Aug 10, 2026. Apply as soon as possible — early applications are often reviewed first.
ML Research Engineer - Post-training (LLMs)
ekacare
You'll be redirected to ekacare's official application page on keka.