Solution Architect (AI/LLM Inference)
Baseten · San Francisco, CA · EPD
About this role
Baseten is hiring a mid-level Solutions Architect in the software engineering function based in San Francisco, CA. The posting calls out experience with Machine Learning, Spark, LLMs, vLLM. Compensation is listed at $165,000–$330,000 per year.
- Role
- Solutions Architect
- Function
- software engineering
- Level
- mid
- Track
- Individual contributor
- Employment
- Full-time
- Location
- San Francisco, CA
- Department
- EPD
- Posted
- May 10, 2026
More roles at Baseten
Job description
from Baseten careersABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $300M Series E, backed by investors including BOND, IVP, Spark Capital, Greylock, and Conviction. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE:
As a Solution Architect (AI/LLM Inference) at Baseten you will partner closely with Sales and customers to translate business needs into technical solutions, run technical discovery, and guide repeatable deployments and proofs of value for customers. This role is a great fit for entrepreneurial, customer-facing technical professionals who want a front-row view into how modern companies adopt AI at scale, and who enjoy working across technical discovery, solution design, demos, deployment scoping, and hands-on customer implementations, in close partnership with Sales and Engineering.
RESPONSIBILITIES:
Partner with Sales on customer discovery calls (most often second calls, occasionally first calls for large accounts).
Lead demos and technical scoping to align on success criteria, architecture, and deployment approach.
This is an excerpt. Read the full job description on Baseten careers →