← Back to sample jobs

Product Engineer - Dedicated Inference

Baseten · series b

📍 San Francisco

fullstackonsite

Listed July 31, 2026

Why this role

Baseten is the model inference infrastructure layer for AI — deploy, scale, and serve any ML model with production-grade GPU optimization, auto-scaling, and enterprise SLAs. Powers inference for leading AI labs and enterprises needing reliable, high-performance model serving.

Sign up free to see the apply link and more roles like this.

Ready to apply?

Get started free → Browse more roles