Machine Learning Engineer
About Elicit
AI research assistant using LLMs over 138M+ papers to automate literature reviews and evidence synthesis for researchers and pharma decision-makers.
Series A · 10–50 people
Job description
The full responsibilities and requirements are on the employer’s site.
Open application page ↗Source and classification
Internal deployment & tooling · Evidence for this classification:
and charts a safer path toward advanced AI tomorrow. Our vision is ambitious: we’re building the default starting point for understanding and reasoning through any hard question. We invite you to help us build that future. (See how people use Elicit today on Twitter; explore our vision in the roadmap.) About the role As a Machine Learning Engineer at Elicit, you’ll build products and workflows that help researchers and scientific teams make higher quality decisions with language models. This is not a role for someone who only wants to develop models in isolation from user impact. A large part of the work is software engineering: building product experiences, APIs, data integrations, evaluation systems, and reliable harnesses that make language models reliably useful and trustworthy in high-stakes domains. You’ll work on problems like: Turning messy, ambiguous research tasks into
More from the job description
About Elicit Elicit is building the reasoning layer for science and decision-making. We use language models to search over 125 million papers, extract data, and surface insights so that researchers, policy-makers, and industry leaders can go from questions to evidence-backed decisions in minutes. Today, hundreds of thousands of researchers have used Elicit to speed up literature reviews, automate systematic reviews, and explore new domains. As we expand our impact beyond academic research, we are laying the groundwork for ML systems that are systematic, transparent, and unbounded when reasoning at scale. To do this, Elicit is pioneering supervision of process, not outcomes. Instead of favoring large black-box models, we break complex questions down into human-legible steps and supervise the reasoning process itself. This approach delivers more transparent, defensible answers today and charts a safer path toward advanced AI tomorrow. Our vision is ambitious: we’re building the default starting point for understanding and reasoning through any hard question. We invite you to help us build that future. (See how people use Elicit today on Twitter; explore our vision in the roadmap.) About the role As a Machine Learning Engineer at Elicit, you’ll build products and workflows that help researchers and scientific teams make higher quality decisions with language models. This i [... source excerpt omitted ...] t assessment, evidence synthesis, and experiment planning that allow models to provide guarantees about their processes Data integrations across literature, scientific databases, customer data, and internal tools APIs that customers can use in their own systems Evaluation systems that help us understand whether a change actually improves user outcomes Trust and transparency features, like source-quality signals, intermediate reasoning, and better ways to inspect and fix outputs Example projects Examples of projects you could work on: Build a target-assessment workflow that combines literature, genetics, chemistry, clinical, regulatory, and company data into a shareable art [... source excerpt omitted ...] d evidence-monitoring workflows that keep teams up to date through alerts, briefs, and living reports. Build enterprise APIs and structured-output pipelines that plug Elicit into customers’ internal systems. Build interfaces that make it easier to inspect, trust, and correct model outputs. Build workflow-specific evals and quality systems that tell us whether a product change actually helped users. Improve extraction, reasoning, or search quality with better prompts, better system design, or finetuning when appropriate. What you bring A strong software engineering background and can build end-to-end systems, not just scripts or notebooks Fluency with language models to rea
Employer postings · Data from · Sources