Senior Manager, AI Inference
Red Hat · Boston, MA
About this role
Red Hat is hiring a senior-level Technical Lead in the software engineering function based in Boston, MA. The posting calls out experience with Kubernetes, Git, Linux, Python. Compensation is listed at $210,730–$358,070 per year.
- Role
- Technical Lead
- Function
- software engineering
- Level
- senior
- Track
- hybrid
- Employment
- Full-time
- Location
- Boston, MA
- Posted
- May 19, 2026
More roles at Red Hat
Job description
from Red Hat careersAbout the Job
At Red Hat we believe the future of AI is open and we are on a mission to bring the power of open-source LLMs and vLLM to every enterprise. Red Hat Inference engineering team accelerates AI for the enterprise and brings operational simplicity to GenAI deployments. As leading maintainers of the vLLM and LLM-D projects, and inventors of state-of-the-art techniques for model quantization and sparsification, our team provides a stable platform for enterprises to build, optimize, and scale LLM deployments.
As a Senior Engineering Manager of the Machine Learning Engineering team focused on vLLM inference, you will be at the forefront of innovation, collaborating with our team to tackle the most pressing challenges in model performance and efficiency. Your technical and people leadership with machine learning and high performance computing will directly impact the development of our cutting-edge software platform, helping to shape the future of AI deployment and utilization. You would be joining the core team behind 2025's most popular open source project on GitHub. If you are someone who wants to contribute to solving challenging technical problems at the forefront of deep learning in the open source way, this is the role for you.