Senior Performance Engineer – AI Platforms

Red Hat • Boston, MA, United States • Posted June 06, 2026

Location Boston, MA
Job Type Full-time
Category other-general
Posted June 06, 2026
**About the Job**

The Red Hat Performance and Scale Engineering team is seeking a Senior Performance Engineer to join our PSAP (Performance and Scale for AI Platforms) team. In this role, you will drive the performance and scalability of distributed inference for Large Language Models (LLMs) as part of the Red Hat AI Inference Server (RHAIIS) open-source project. You will be responsible for characterizing, modeling, and understanding performance deltas to ensure industry-leading throughput, latency, and cost-efficiency of AI workloads. This includes using tools like vLLM, GuideLLM, and PyTorch for example.This is a dynamic role for a seasoned engineer with a growth mindset who handles and adapts to rapid change, has a strong commitment to open-source values, and the willingness to learn and apply new technologies. You will be joining a vibrant open source culture and helping promote performance and innovation in this Red Hat engineering team.

The broader mission of th...

Interested in this role?

Click the button below to start your application.

Apply Now