Design and optimize high-performance GPU and accelerator kernels for a high-throughput AI inference platform. Identify performance bottlenecks in production LLM workloads and implement hardware-aware optimizations. Requires deep expertise in CUDA, C++, and modern GPU architectures to improve model latency and throughput.
Key Highlights
Searching for Development & Programming roles that provide visa sponsorship? Connect with international employers through Development & Programming Jobs with Visa Sponsorship opportunities actively seeking talented professionals.
Key Responsibilities
Technical Skills Required
Explore our comprehensive directory of visa sponsorship jobs from employers worldwide who are ready to sponsor talented international professionals.
Benefits & Perks
Interested in opportunities specifically in United State? Discover our dedicated Visa Sponsorship Jobs in United State page featuring roles from top employers in this location.
Nice to Have
Job Description
Member of Technical Staff - Kernel Engineer
Location: Santa Clara, CA
About the Company
This company builds and operates a high-performance AI inference platform that helps enterprises run large language models faster, more efficiently, and at scale.
Its infrastructure sits at the foundation of the AI application stack, with a focus on improving the speed, reliability, utilization, and economics of production model serving.
About the Role
A fast-growing AI infrastructure company operates an enterprise inference platform providing high-throughput, low-latency access to large language models across a distributed, heterogeneous compute fleet.
This is a short excerpt. All rights to the full description belong to its original publisher.
Similar Jobs
Explore other opportunities that match your interests