This position is no longer accepting applications
Closed on August 10, 2026.
This role is filled — get an email when new Software Engineering roles open on DeveloperJobs.io:
Principal Machine Learning Engineer
Ai Ml
Artificial Intelligence
Cloud Operations
DevOps
Engineering
Large Language Models
Machine Learning
Ml Ops
Performance Engineering
Software Engineering
Technical Lead
View similar jobs
Get alerted when similar jobs are posted — set up a New Software Engineering jobs on DeveloperJobs.io alert.
See other roles at Oracle.
Job Description
Oracle's Generative AI Services team is seeking a Principal Machine Learning Engineer to shape the architecture, design, and deployment of distributed AI infrastructure. This onsite role in Seattle, WA focuses on building scalable systems for training, fine-tuning, and inference of large models, coordinating with partner engineering teams to deliver reliable production deployments and contributing to open-source frameworks such as vLLM and SGLang.
Responsibilities
- Lead the architecture, design, and development of distributed, scalable, high-performance systems for AI model training, fine-tuning, and inference.
- Build and optimize next-generation AI infrastructure that powers large-scale generative workloads.
- Direct analysis of model architectures to improve performance, efficiency, and scalability.
- Utilize cutting-edge technologies to craft state-of-the-art AI systems and onboard frontier models.
- Benchmark, diagnose, troubleshoot, and resolve issues across the AI model lifecycle, including training, fine-tuning, and serving, ensuring reliable production deployments.
- Contribute to open-source frameworks such as vLLM and SGLang and strengthen OCI's involvement.
- Oversee a team of senior and junior engineers, guiding delivery of the roadmap on schedule with high quality.
Technologies
- vLLM
- SGLang