This one is closed
Live roles like this one
- C 11h ago
-
B
10h ago
Upstart Anywhere in the World Salary Range $141,000
-
D
18h ago
Principal Data Scientist, Generative Recommendations & Agentic Orchestration- Sp
Cimpress/Vista Spain
- C 1d ago
See every "Data Scientist" role →
Get new “Data Scientist” roles by email
One email a day with what is new in "Data Scientist". Nothing new, no email.
We confirm the address first, and every mail carries an unsubscribe link. Alerts are ours, not a third party's.
Why this grade This listing scored 48/100, which is a D. It lost the most ground on pay transparency. See the breakdown
- Description depth 20 / 20 How much the posting actually says about the work, measured in characters of real text.
- Pay transparency 12 / 25 A published salary range, worth more than any other single factor because it is what a candidate cannot find out without applying.
- Freshness 8 / 15 How recently it was posted. Older postings are likelier to be filled or abandoned.
- Remote clarity 8 / 15 Whether "remote" means anywhere, or is quietly restricted to one country.
- Corroboration 5 / 10 Whether more than one source carries this listing.
- Role specificity 0 / 10 Whether the listing is tagged well enough to tell what the role actually is.
-5 Ghost-job penalty — Deducted for signals that this posting may not be a real, currently-open role — staleness, repeated relisting, or talent-pool language.
Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →
This listing does not state a salary
$123k – $169k
That is the middle half of what comparable roles paid on this board over the last 90 days — 15 listings that did publish a figure, median $154k. It is not this employer's offer, and we have no idea what they pay. It is only what the rest of the market advertised.
Your Future Evolves Here
Evolent partners with health plans and providers to achieve better outcomes for people with most complex and costly health conditions. Working across specialties and primary care, we seek to connect the pieces of fragmented health care system and ensure people get the same level of care and compassion we would want for our loved ones.
Evolent employees enjoy work/life balance, the flexibility to suit their work to their lives, and autonomy they need to get things done. We believe that people do their best work when they're supported to live their best lives, and when they feel welcome to bring their whole selves to work. That's one reason why diversity and inclusion are core to our business.
Join Evolent for the mission. Stay for the culture.
What You’ll Be Doing:
Core Responsibilities
- Own the diagnostic loop for LLM-based clinical services: take failure modes surfaced by clinical reviewers and product managers, form root-cause hypotheses (prompt design, context assembly, retrieval, guideline encoding, model behavior, upstream data), and design experiments that isolate the cause.
- Test candidate fixes — prompt and configuration variants, context changes, model alternatives — and verify improvements with structured evaluations, not anecdotes; confirm fixes don't regress other behavior.
- Own the evaluation roadmap and quality bar for the team's AI services: which metrics gate deployment, how golden sets and regression suites grow, and what "good enough to ship" means in evidence.
- Use and extend the team's evaluation platform: build golden sets and regression suites, define metrics (accuracy, guideline adherence, grounding/faithfulness), run and interpret eval batteries. Fluency in *using* modern eval tooling matters; the platform exists — extending it thoughtfully is in scope, rebuilding it is not.
- Partner with clinical reviewers (medical directors) to turn review findings into labeled evidence and executable evaluation criteria; partner with product managers to prioritize which failure modes matter most.
- Mentor others in evaluation methods and grow the evals function as it scales, including readiness to take direct reports as the team expands.
- Support the annual clinical-guidelines update cycle with regression evaluation as guidelines, prompts, and models change; ramp with the team's senior data scientists in Q4 2026.
- Document the diagnostic playbook: failure taxonomies, experiment templates, variant history — a method others can run, not a private intuition. This is a practicing role — the diagnostic loop is the job, at every level of seniority.
- Handle clinical data (including PHI) according to organizational security, privacy, and compliance requirements.
Minimum Requirements
- Bachelor's degree in Data Science, Computer Science, Statistics, or a related quantitative field — or equivalent experience.
- 5+ years of data science or applied machine learning experience, including shipping and maintaining models or AI systems in production.
- 2+ years of recent, hands-on experience evaluating and improving LLM-based systems: structured error analysis, prompt/configuration iteration, experiment design, metrics interpretation — as a practitioner, not only as a reviewer of others' work.
- Demonstrated experience with LLM evaluation methods and tools — golden/regression sets, LLM-as-judge with validation, tracing and observability tooling.
- Strong Python and solid data-analysis skills (SQL a plus); comfort computing and reasoning about metrics such as sensitivity, specificity, and PPV.
- Hypothesis-driven working style: the instinct to isolate variables and prove a fix, rather than tweak and hope.
- Strong written communication — findings and go/no-go evidence must be legible to engineers, clinicians, and leadership.
Preferred Requirements
- Healthcare experience: utilization management, prior authorization, clinical documentation, or clinical/claims data; comfort reading clinical guideline and medical-policy content.
- Experience mentoring data scientists or leading small technical workstreams; interest in growing into people leadership as a function scales.
- Experience working with clinical reviewers or other domain experts to convert expert judgment into labeled data and evaluation criteria.
- Experience with LLM observability/telemetry stacks (OpenTelemetry-based tracing, Logfire, Langfuse, or similar).
- Statistics or experimentation background (A/B testing, statistical significance, sample-size reasoning).
- Master's degree in a quantitative field.
To ensure a secure hiring process we have implemented several identity verification steps, including submission of a government issued photo ID. We conduct identity verification during interviews, and final interviews may require onsite attendance. All candidates must complete a comprehensive background check, in-person I-9 verification, and may be subject to drug screening prior to employment. The use of artificial intelligence tools during interviews is prohibited and monitored. Misrepresentation will result in immediate disqualification from consideration.
Technical Requirements:
We require that all employees have the following technical capability at their home: High speed internet over 10 Mbps and, specifically for all call center employees, the ability to plug in directly to the home internet router.
Evolent is an equal opportunity employer and considers all qualified applicants equally without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, or disability status.If you need reasonable accommodation to access the information provided on this website, please contact for further assistance.
The expected base salary/wage range for this position is $135,000 - 165,000. As part of our total compensation package, Evolent is proud to offer comprehensive benefits (including health insurance benefits) to qualifying employees. All compensation determinations are based on the skills and experience required for the position and commensurate with experience of selected individuals, which may vary above and below the stated amounts.Originally posted on Himalayas
Apply for this role Opens himalayas.app — the link as listed; we have not yet verified it is the employer's own page
Quick question · anonymous · one tap
Would you apply to this job?
Answer to see what other job seekers said.
Your turn · no account needed
Help the next applicant
You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.
I know what it pays
Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.
Where this listing came from
- 31 Aug 2026 Himalayas first sighting
Seen on 1 board over 0 days.