Smarter Dx

Staff Machine Learning Research Scientist

Smarter Dx · Remote (US)
Remote (US) Remote was $220K–$260K Closed
Applications are closed for this role. It was originally posted 2026-07-16. It’s no longer accepting applicants — see roles Smarter Dx is still hiring for →, or browse the live openings below.
Salary
$220K–$260K
Type
Full-time
Experience
8+ yr

SmarterDx is transforming how health systems use clinical AI to capture the full value of patient care delivered. Built by physician-data scientists and trained on clinically-validated EHR data, our clinical AI platform interprets the nuances behind every patient story and makes clinically-sound recommendations for revenue cycle teams — helping hospitals recover earned revenue, improve quality metrics, reduce denials, and streamline revenue cycle operations. As a Smartian, you’ll help build technology that makes healthcare more accurate, sustainable, and effective for everyone. Learn more at smarterdx.com/careers .

As a Staff Machine Learning Research Scientist at SmarterDx, you will set technical direction for cutting-edge ML research and translate it into real-world clinical impact. You’ll work at the intersection of research, engineering, and healthcare, partnering with engineers and clinicians to build systems that deeply understand patient records and improve hospital outcomes. This is a senior, high-impact role where you’ll not only execute on ambitious ideas but also shape the team’s research agenda and standards.

You will be expected to operate with a high degree of autonomy—identifying promising research directions, critically evaluating academic work, and ensuring that what gets built is both scientifically sound and practically useful. Your work will directly influence how we evaluate models, detect hallucinations, and build high-quality datasets, ultimately improving the reliability of AI in healthcare.

  • *This role is fully remote within the US**

What You’ll Do

  • Lead end-to-end ML research, from idea generation to production deployment and monitoring
  • Design, implement, and evaluate novel methods for LLM alignment on proprietary clinical data
  • Develop and rigorously evaluate approaches for hallucination detection, attribution, and model reliability
  • Build and curate high-quality datasets, with a strong emphasis on evaluation design and benchmark integrity
  • Critically assess academic literature to identify strong vs weak methods, and translate the best ideas into practice
  • Establish best practices for experimental design, including statistically sound evaluation and reproducibility
  • Collaborate cross-functionally with engineering to productionize models (MLOps, infra, deployment)
  • Develop methods for long-context and multimodal modeling (structured + unstructured clinical data)
  • Mentor other researchers and help raise the bar for research quality across the team
  • Contribute to external presence through papers, talks, and recruiting

What You Bring

  • Strong track record of ML research, ideally with publications in top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP, AAAI, etc.)
  • Proven ability to distinguish high-quality vs low-quality research, especially in fast-moving areas like LLMs
  • Deep understanding of LLM failure modes, particularly hallucinations, and how to evaluate and mitigate them
  • Experience designing rigorous evaluation frameworks and building high-quality test datasets
  • Strong intuition for dataset quality, bias, and benchmark design (data-centric AI mindset)
  • Hands-on experience training large-scale deep learning models (multi-GPU / distributed systems)
  • Deep understanding of modern neural architectures (transformers, SSMs, encoder/decoder models, etc.)
  • Strong programming skills in Python and ML frameworks such as PyTorch or JAX
  • Experience deploying ML models into production systems and monitoring their performance
  • Clear and proactive communicator, able to explain complex ideas and critique work effectively

Nice To Haves

  • Experience with inference optimization techniques (e.g., vLLM, KV caching, speculative decoding)
  • Familiarity with MLSys concepts (parallelism strategies, distributed training infrastructure)
  • Experience working with clinical or healthcare data
  • Background in retrieval systems, graph-based learning, or multimodal modeling

Our Tech Stack

  • PyTorch, Hugging Face Transformers, Python
  • AWS (MWAA), Kubernetes, SLURM
  • DeepSpeed, TorchTune
  • Snowflake, Airflow, GitHub

Compensation

$220k-260k base salary

#LI-Remote

Benefits

  • Medical, Dental & Vision – Comprehensive plans with leading insurance providers, covering 75% of your premiums, depending on the plan.
  • Paid Parental Leave – Generous paid leave to support families through birth or adoption: Up to 12 weeks for parents.
  • Remote-First Team – Work from anywhere in the U.S.
  • Unlimited PTO & 10 Holidays – So you can relax and recharge.
  • 401(k) with Traditional & Roth Options  – Tax-advantaged retirement savings through Fidelity with a 4% match.
  • Minimal Bureaucracy  – A fast-moving, high-impact environment where you can focus on what matters.
  • Incredible Teammates!  – Work alongside smart, supportive, and mission-driven colleagues.
LLMPythonPyTorchAWSKubernetesSnowflake
E
Staff Security Engineer
Remote (US) Remote
Engineering
$230K–$250K
E
Staff Software Engineer, Machine Learning
Remote (US) Remote
Engineering
$230K–$250K
E
Senior Software Engineer, Backend
Remote (US) Remote
Engineering
$190K–$230K
See all live roles at Smarter Dx →
G
Staff Applied Scientist
Garner Health Technology New York, NY Hybrid
Data & ML
$300K–$390K
S
Research Scientist - RL Training
Snorkel AI Redwood City, CA Remote
Data & ML
$200K–$350K
B
Staff Machine Learning Engineer, Underwriting and Credit
Block Bay Area, CA, United States of America
Data & ML
$276K–$415K
R
Senior Staff Machine Learning Engineer, GenAI Platform
Reddit Remote (US) Remote
Data & ML
$292K–$409K
See all Data & ML roles →
Applications closed