U

Senior Software Engineer (C++ Systems) – GPU Virtualization & Performance Optimization

USA Tech Recruit San Francisco Bay Area
Visa Sponsorship Relocation
Apply
AI Summary

This role focuses on optimizing high-performance C++ virtualization libraries for GPU systems, pushing microsecond-level performance boundaries. You’ll work on advanced topics like oversubscription, checkpointing, and distributed GPU clusters while supporting new hardware architectures. Ideal for engineers with deep systems expertise and a passion for low-latency, production-grade infrastructure.

Key Highlights
Optimize high-performance C++ virtualization library for ultra-low-latency execution
Investigate and implement advanced GPU virtualization techniques (oversubscription, checkpointing, distributed clusters)
Debug and diagnose performance bottlenecks in production ML workflows and cross-stack systems
Key Responsibilities
Optimize a high-performance C++ virtualization library for low-latency execution
Investigate and implement advanced GPU virtualization techniques such as oversubscription, checkpointing, and distributed GPU clusters
Support new hardware architectures with deep cross-stack systems insight
Perform systems-level debugging in live production environments
Diagnose and resolve performance issues in machine-learning workflows across the stack
Technical Skills Required
C++ Systems Programming Performance Optimization
Benefits & Perks
Visa sponsorship for TN, Canadian, and H-1B transfers (not new applications)
Relocation package available
Onsite position in San Francisco
Nice to Have
Rust expertise
Experience with compilers
Background in optimizing NIC or C++ performance

Job Description


We're partnering with a high-growth client in the GPU and systems-software space to help them hire a Software Engineer (C++ Systems). This role sits at the heart of their GPU virtualization technology. It’s a strong fit for someone who enjoys pushing microsecond-level performance boundaries in complex C++ systems and wants to make an impact on low-level GPU infrastructure used at scale.


This would be a full time, onsite position in SF, and they are able to provide TN / Canadian / H-1B visa transfers but not new applications at this time. Relocation packages are also available.


Key Responsibilities:

  • Optimize a high-performance C++ virtualization library for low-latency execution.
  • Investigate advanced topics such as oversubscription, checkpointing, and distributed GPU clusters.
  • Support new hardware architectures with deep, cross-stack systems insight.
  • Perform systems-level debugging in live production environments.
  • Diagnose performance issues in machine-learning workflows across the stack.


Key Qualifications:

  • Exceptional C++ proficiency (Rust expertise also considered, though day-to-day work is in C++).
  • Proven experience building and maintaining low-level systems in production.
  • Strong academic background in Computer Science or a related field (high GPA preferred).
  • Experience with compilers, networking protocols, or kernel-level development.
  • Background in optimizing NIC or C++ performance.
  • Ability to trace and resolve performance bottlenecks through multiple layers of a system.


By applying to this role you understand that we may collect your personal data and store and process it on our systems. For more information please see our Privacy Notice https://eu-recruit.com/wp-content/uploads/2024/07/European-Tech-Recruit-Privacy-Notice-2024.pdf


Similar Jobs

Explore other opportunities that match your interests

Forward Deployed Engineer

Programming
5h ago
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Associate

benchstack ai

San Francisco Bay Area
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

retell

San Francisco Bay Area
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Mid-Senior level

Stealth Startup

San Francisco Bay Area

Subscribe our newsletter

New Things Will Always Update Regularly