EU remote
4492 - Lead Data Scientist
About this role
Your Role Hundreds of billions of dollars are spent every year in American healthcare on administrative work that never reaches a patient. That is not a technology gap. It is an operating model that answered fragmentation by adding more people, more processes, and more systems. The limiting factor in changing that is no longer model quality. Frontier models are widely available and improving monthly. What decides whether an AI system actually holds up in a hospital or a health plan is everything around the model: the context it inherits, the retrieval it depends on, the evaluation that proves it works, and the guardrails that keep it inside its authority.
Demos are cheap. Systems that survive real clinical, financial, and audit conditions are not. Anyone can demo an agent. The scarce skill is proving it still works on the ten-thousandth chart, at a cost per case the business can defend. This role sits exactly there. You will own a defined slice of the Applied AI roadmap — architecting the agentic and retrieval systems behind it, owning their lifecycle in production, and growing a small team of engineers who build alongside you.
It is deliberately a hands-on leadership role: you keep writing and reviewing code while owning delivery, quality, and the growth of the people on your team. Analytics at Innovaccer Healthcare generates more data than almost any other industry, yet most of it stays siloed across disconnected platforms. Innovaccer's autonomous operations platform unifies that data and the Analytics team closes that gap between data and decisions.
They build production ML systems that turn fragmented data into actionable decisions: risk scores that flag deteriorating patients, operational models that surface inefficiencies, and prescriptive tools that guide next steps. Analytics is core product infrastructure, and this team sets the standard for what that means in practice. If that is what you are looking for, this is the team. A Day in the Life You own AI systems end to end — from problem framing with product and customer teams through architecture, evaluation, deployment, and the cost and latency profile they run at.
System ownership · Architecture. Design production agentic and RAG systems that meet real customer scale and reliability requirements, not benchmark conditions. · Orchestration depth. Apply agent design patterns with judgment — memory, tool routing, multi-agent coordination, and explicit failure handling — and decide deliberately where autonomy stops and a human takes over. · Retrieval quality. Own the retrieval pipeline as a first-class system: chunking strategy, embedding model selection, vector stores, re-ranking, and relevance tuning against measured outcomes.
· LLMOps lifecycle. Own model and prompt versioning, eval pipelines running in CI, observability (tracing, token and cost dashboards), and the guardrails and safety filters that ship with every release. · Model strategy. Select, fine-tune, and serve SLMs and open models — LoRA and QLoRA, quantization, inference optimization, GPU and serving trade-offs. · Engineering economics. Optimize latency, cost, and accuracy together, and make defensible build-versus-buy and model-selection calls you can explain to both engineers and executives.
Beyond the codebase · Define and execute the quarterly roadmap for your area, and translate ambiguous business problems into machine learning problems with clear solution workflows. · Work with business leaders and customers directly to understand where the workflow actually breaks, then build for that. · Partner with the data platform and applications teams so your capabilities land inside their products and workflows rather than beside them.
· Set the coding and evaluation standards for your team, and lead design reviews. · Pursue published work or patents where the problem warrants it — particularly in healthcare AI. What You Need · 6+ years in data science, applied ML, or AI engineering, including 2+ years building LLM-powered products. Healthcare experience is a plus. · Deep NLP and GenAI experience. Statistical and classical machine learning is good to have on top of that.
· Strong hands-on Python — building highly scalable, performant enterprise applications, plus optimization technique. · Hands-on experience with deep learning frameworks: PyTorch and/or HuggingFace transformers. · At least one shipped GenAI product with a genuinely complex architecture — multiple agents, memory, retrieval, and agent OTEL/tracing in production. · Working command of modern fine-tuning: PEFT methods, with LoRA and QLoRA preferred.
· Hands-on experience with at least one ML platform — Databricks, Azure ML, or SageMaker. · Experience leading engineers, formally or as a technical lead — you have owned other people’s output, not only your own. · Strong written and spoken communication, with a customer-focused instinct in both conversation and documentation. · Preferably a Master’s in Computer Science, Computer Engineering, or a related field. The engineering baseline we assume Everything above sits on top of independent delivery.
This role assumes you can already do the following without supervision: · Build production-grade RAG and LLM/SLM-powered features end to end with limited supervision. · Work fluently in at least one orchestration framework — LangGraph, LlamaIndex, CrewAI, or equivalent — to compose multi-step, tool-using flows. · Design retrieval pipelines and tune them for relevance: chunking, embeddings, vector stores, re-ranking. · Implement prompt engineering, function and tool calling, and reliable structured-output parsing.
· Write and run evals — golden sets, LLM-as-judge — to measure quality and catch regressions before customers do. · Containerize and deploy services (Docker, REST/gRPC) with an eye on latency, token cost, and basic guardrails. · Document well and participate actively in code review. Good to have · API frameworks for robust web applications — FastAPI or Django preferred. · Comfort with at least one hyperscaler cloud. · Papers or patents, especially in healthcare AI.
We offer competitive benefits to set you up for success in and outside of work. Here’s What We Offer Generous Leave Benefits: Enjoy generous leave benefits of up to 40 days. Parental Leave: Experience one of the industry's best parental leave policies to spend time with your new addition. Sabbatical Leave Policy: Want to focus on skill development, pursue an academic career, or just take a break? We've got you covered.
Health Insurance: We offer health benefits and insurance to you and your family for medically related expenses related to illness, disease, or injury. Pet-Friendly Office*: Spend more time with your treasured friends, even when you're away from home. Bring your furry friends with you to the office and le
Source listing: workable_innovaccer-analytics