$165K to $330K a year as published by the employer
Applications are completed on the employer's own site.
Oversee execution systems and cross-functional engineering programs for AI inference platform model releases and optimization.
# Technical Program Manager, Model Performance
Baseten
Technical Program Manager building execution systems for Baseten’s high-performance AI inference platform. Coordinating model releases, optimization, and cross-functional engineering programs.
Posted 8/29/2026full-timeSan Francisco • California • 🇺🇸 United StatesMid-LevelSenior💰 $165,000 - $330,000 per yearWebsite
## Core Competencies
Role fit
Core Competencies
Use this summary to align your resume positioning with the role.
Demonstrates deep technical program management expertise with a focus on model performance and inference optimization. Capable of driving cross-team alignment and managing complex projects while maintaining clear communication and accountability.
Highest-signal resume keywords
Technical Program ManagementModel Performance OptimizationCross-Team AlignmentInfluencing Without AuthorityHigh-Agency Decision Making
## ATS Keywords
Tailor your resume
Applicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Program ManagementModel OptimizationProject ExecutionRisk ManagementStatus Reporting
Soft Skills
Excellent CommunicationComfort with AmbiguityOwnershipAccountability
Tools & Technologies
VLLMTensorRT-LLMSGLangNVIDIA Dynamo
Industry Keywords
Model PerformanceInference Production StackPerformance EngineeringProject Portfolio Management
### About the role
Key responsibilities & impact
* Own execution across Model Performance's active project portfolio * Design and establish planning structures, operating cadences, and status reporting mechanisms * Coordinate model release and optimization programs end to end, including day-zero launches * Sequence work across performance engineering, infrastructure, and release stakeholders * Drive cross-team alignment as scope expands from Model Performance Core into Model APIs and the inference production stack (BIS) * Surface risks and dependencies early and keep leadership informed with clear, honest status * Partner with engineering leads to design team structures and ownership boundaries as the organization scales
### Requirements
What you’ll need
* Deep technical program management experience; already running programs of this scope at an organization of similar or greater complexity * Experience program-managing model performance or inference optimization work * Understanding of how vLLM, TensorRT-LLM, SGLang, or NVIDIA Dynamo fit into a production serving stack * Comfort with ambiguity and zero-to-one program building * Ability to influence without authority across engineers, managers, and leadership * Excellent written and verbal communication * High-agency decision making demonstrating ownership, accountability, and a strong desire to get things done
### Benefits
Comp & perks
* Competitive compensation, including meaningful equity * 100% coverage of medical, dental, and vision insurance for employee and dependents * Flexible PTO policy including company wide Winter Break * Paid parental leave * Fertility and family-building stipend through Carrot * Company-facilitated 401(k) * Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities
Jobtailor is a recruiting intermediary entity.