Member of Technical Staff - Software Engineer, AI Capabilities

preference model • Canada

Visa Sponsorship Relocation

Apply

AI Summary

Preference Model is seeking a Software Engineer to build complex RL training environments for AI models. This role involves designing, coding, and refining environments that expose AI failures in software engineering tasks. Requires deep software engineering expertise, Python proficiency, and experience with coding agents.

Key Highlights

Builds software engineering environments to train AI models.

Owns complex tasks end-to-end with full autonomy.

Blends research and engineering with a focus on novel approaches.

Key Responsibilities

Design, build, and refine RL tasks across their full lifecycle.

Own the hardest environments on the roadmap: multi-step workflows, realistic stakeholder interactions, large codebases with real conventions and technical debt, and system design problems.

Direct coding agents heavily in your day-to-day work, evaluate their output critically, and recognize when they are failing in subtle ways.

Distinguish genuine model capability gaps from grader or environment issues, and redesign tasks to target deeper, more subtle failure modes.

Contribute to the shared infrastructure and tooling that the environments team depends on.

Mentor newer engineers on the team as it grows.

Technical Skills Required

Python

Benefits & Perks

Competitive cash and equity compensation

Health, vision, dental, benefits

401K match

Lunch provided everyday onsite

Weekly snack orders

Visa sponsorship & relocation support available

Nice to Have

Deep specialty expertise in an area that current models struggle with (distributed systems, low-level performance, security, compilers)

Been an early engineer at a previous startup, shipped independently, and want to do it again in AI.

Spent significant time building with coding agents, written about their failure modes, or contributed to agent evaluation work.

Job Description

About Us

Preference Model is building automated ML research engineering.

Existing frontier models are brittle when applied to real-world ML tasks. The present bottleneck is the lack of high-quality RL training environments. Our first step is to build RL environments that reflect real-world complexity, with diverse tasks and robust reward functions.

Our founding team has previous experience on Anthropic’s data team building data infrastructure, and datasets behind Claude. We are partnering with leading AI labs to push AI closer to achieving its transformative potential.

About The Role

AI models have gotten good at narrow coding tasks but still fail at the complex, judgment-heavy parts of software engineering: working in a large codebase with real conventions and technical debt, making the right tradeoff on a system design problem, or navigating a multi-step task with ambiguous stakeholders. As a Member of Technical Staff - Software Engineer on the Capabilities team, you will build software engineering environments that expose those failures and help models improve on them.

This role blends research and engineering. It will require you to stay up to date with the latest research, develop novel approaches, and realize them in code. You will have full ownership and autonomy of the environments you build.

You will own our most complex tasks end-to-end: environments with multi-step workflows, realistic stakeholder interactions, large codebases, and challenging system design problems. You will work closely with a small team of engineers and directly with our founders, and you will ship environments that go into the training loops of frontier models at our partner labs. This is independent, high-ownership work with regular feedback.

What You Will Do

Design, build, and refine RL tasks across their full lifecycle, from ideation through grading, failure analysis, and iteration.
Own the hardest environments on the roadmap: multi-step workflows, realistic stakeholder interactions, large codebases with real conventions and technical debt, and system design problems.
Direct coding agents heavily in your day-to-day work, evaluate their output critically, and recognize when they are failing in subtle ways.

Looking to advance your Development & Programming career with relocation support? Explore Development & Programming Jobs with Relocation Packages that include comprehensive packages to help you move and settle in your new role.

Distinguish genuine model capability gaps from grader or environment issues, and redesign tasks to target deeper, more subtle failure modes.
Contribute to the shared infrastructure and tooling that the environments team depends on.
Mentor newer engineers on the team as it grows

What We are Looking For

Deep software engineering experience across multiple domains, with genuine expertise in at least one specialty: infrastructure, distributed systems, performance, security, compilers, databases, or similar.
Proficiency in Python.
Extensive hands-on experience with coding agents (Claude Code, Cursor, Codex, or similar), including an intuition for where they cut corners and how to direct them well.
Strong intuition for how models behave, even without prior ML or AI experience. You can anticipate where a model will take shortcuts and design around that.
Comfort working independently on complex, ambiguous problems with minimal direction.

Discover our full range of relocation jobs with comprehensive support packages to help you relocate and settle in your new location.

Track record of owning work end-to-end in previous roles.

You may be a good fit if one of the following applies

You have been a senior or staff engineer at a company known for engineering rigor (e.g., a frontier lab, infrastructure startup, or systems-heavy team) and want to apply that experience to model training.
You have deep specialty expertise in an area that current models struggle with (distributed systems, low-level performance, security, compilers) and can design tasks that expose those weaknesses.
You have been an early engineer at a previous startup, shipped independently, and want to do it again in AI.
You have spent significant time building with coding agents, written about their failure modes, or contributed to agent evaluation work.

What we offer:

Interested in relocating to Canada? Check out our comprehensive Relocation Jobs in Canada page with detailed relocation packages and benefits.

Competitive cash and equity compensation (>90th percentile)
Ownership and autonomy in a fast moving startup environment
Opportunity to work with top machine learning engineers
Health, vision, dental, benefits
401K match
Lunch provided everyday onsite
Weekly snack orders
Visa sponsorship & relocation support available

We value diverse perspectives and experiences. If you're excited about this role but don't check every box, we still encourage you to apply.

Compensation Range: $180K - $300K

Job Overview

Posted Date Jun 14, 2026

Employment Type Full-time

Experience Level Not Applicable

Location Canada

Annual Salary 180,000 - 300,000 USD

Category Programming

Company preference model

Mentioned Skills

Industries

Similar Jobs

Explore other opportunities that match your interests

Senior Systems Engineer - Satellite Systems Designer and Functional Lead

Programming

•

1d ago

Premium Job

•••••• •••••• ••••••

Job Type ••••••

Experience Level ••••••

kepler communications inc.

Canada

Frontend Engineer - JavaScript, React, and Web Performance

Programming

•

1d ago

Visa Sponsorship Relocation Remote

Job Type Full-time

Experience Level Not Applicable

Ramp

Canada

Frontend Software Engineer

Programming

•

2d ago

Visa Sponsorship Relocation Remote

Job Type Full-time

Experience Level Not Applicable

Harvey

Canada

Member of Technical Staff - Software Engineer, AI Capabilities

Key Highlights

Key Responsibilities

Technical Skills Required

Benefits & Perks

Nice to Have

Job Description

Job Overview

Mentioned Skills

Industries

Similar Jobs

Senior Systems Engineer - Satellite Systems Designer and Functional Lead

Premium Job

kepler communications inc.

Frontend Engineer - JavaScript, React, and Web Performance

Ramp

Frontend Software Engineer

Harvey

Subscribe our newsletter