Companies
← All boards

Applied AI Engineer

Sully.aiUnited StatesRemote · United States only$100K–$130KEstimated

The posting does not state a salary. This range is our estimate from the role, the location and the stack — treat it as a guide, not an offer.

Level
Mid
Apply from
United States

AI Engineering

Apply at Sully.ai

About the role

Skills the posting asks for

Summarised from the employer’s posting, which is reproduced in full below.

The employer’s full posting

About Sully.ai

Sully.ai is building the AI operating system for healthcare. Our mission is to make doctors superhuman through AI, so every patient can get faster, better, and more accessible care.

Healthcare is one of the most important systems in the world, yet clinicians still lose hours every day to documentation, scheduling, intake, coding, and administrative work. We are changing that with AI medical employees: Scribes, Receptionists, Coders, Consultants and other agents that operate inside real clinical workflows.

We are backed by Amity Ventures, YC, Baidu Ventures amongst others, have raised $40M+, and are already deployed across hundreds of healthcare organizations. Our products process tens of thousands of clinical encounters every day, automate millions of minutes of administrative work, and help clinicians spend more time with patients.

The role

As an Applied AI Engineer on the Research team, you will work directly with our research and engineering org on the AI medical employees at the core of Sully — Scribe, Receptionist, Consultant, Coder, and what comes next. Your job is to turn the frontier of agent research into agents and eval systems that run in production for real clinicians within days.

This is a builder-researcher role. You read the latest work in agent, harness and context engineering, evals, and reinforcement learning, then ship from it. We do not separate "research" and "engineering" — you are expected to do both, and the work is measured by what lands in front of clinicians, not by a benchmark in isolation. Our agents are architecture- and configuration-driven: most of the behavior comes from decomposition, prompts, routing logic, context assembly, and evals rather than one-off code. You will own the systems that decide what the model sees, measure whether it worked, and make it better on its own over time.

You will own three things: evals, agents, and self-improving loops. Just as importantly, you will help define how Sully does applied research at all — the harnesses, eval infrastructure, and feedback loops that make every future agent faster to build and safer to ship. The environment is high-stakes by design: in clinical AI, accuracy has zero margin for error, and that constraint is the most interesting part of the job.

What you’ll do

Build and scale evaluation systems

Build and ship medical AI agents

Build self-improving and auto-evolving loops

Push context and loop engineering forward

Partner across research, engineering, and clinical

What good looks like What we’re looking for

Required

Strong signals

Who this role is not for

This role may not be a fit if you:

Interview process

Location: Open to US remote, but ideally office based/hybrid in Bay Area.