Hippocratic AI Inc. is seeking an LLM Inference Engineer to own the serving infrastructure enabling fast, reliable healthcare AI for patient conversations. You will optimize end-to-end latency ( By day 90 you will ship measurable improvements to the inference stack, validate gains, and outline a performance optimization plan. At 12 months you will deploy advanced serving architectures and contribute techniques that become core infra capabilities.

Also on the board Same function, level within a rung

Level

Lead

Location

Menlo Park, CA

Occupation

Computer Systems Engineers/Architects

Industry

All Other Miscellaneous Ambulatory Health Care Services

Posted

yesterday

Apply for this role →