Software Engineer - ML Systems & AI Infrastructure
Build and optimize production infrastructure for LLM model compression and inference. Develop high-performance ML systems focusing on quantization, pruning, and GPU execution. Requires strong software engineering fundamentals and deep expertise in GPU performance and PyTorch.
Key Highlights
Key Responsibilities
Looking to advance your Development & Programming career with relocation support? Explore Development & Programming Jobs with Relocation Packages that include comprehensive packages to help you move and settle in your new role.
Technical Skills Required
Benefits & Perks
Discover our full range of relocation jobs with comprehensive support packages to help you relocate and settle in your new location.
Nice to Have
Job Description
At Orient Path, we partner with high-growth technology companies globally.
One of our clients is an early-stage deep-tech AI company developing a fundamentally new approach to LLM model compression — making foundation models smaller, faster, and more efficient to run in production.
Their research team develops compression methods. Now they are building the engineering layer that turns this research into reliable, high-performance ML infrastructure.
What are we building?
Production infrastructure around LLM compression and inference.
This is a short excerpt. All rights to the full description belong to its original publisher.
Similar Jobs
Explore other opportunities that match your interests