[Skip To Content]

Senior Manager, AI Deployment

  • 위치
    • Remote
    • Sunnyvale, California
  • 직무 유형 Full time
  • 게시됨
  • Job Requisition JR-202619803

설명

About the Organization 

General Motors is developing the software and artificial intelligence capabilities for the next generation of autonomous driving. Within AI Foundations, AI Acceleration makes machine learning models faster, more efficient, and more reliable on production vehicle hardware. 

The AI Deployment team owns model inference performance across simulation, hardware-in-the-loop, bench, and vehicle environments. The team focuses on latency, memory, GPU utilization, numerical parity, profiling, benchmarking, reduced precision, and production readiness. 

About the Role 

We are looking for a Senior Manager, AI Deployment to lead the strategy and execution of model performance and on-vehicle inference for autonomous driving. You will lead engineering managers and senior technical leaders working across model optimization, GPU systems, inference runtimes, and vehicle integration. You will set performance goals, guide optimization of complex autonomy models, and establish disciplined methods to measure latency, diagnose regressions, and validate improvements. Success requires strong technical judgment, people leadership, and the ability to make clear trade-offs among latency, memory, throughput, accuracy, power, and numerical parity. 

What You’ll Do 

  • Own the strategy, roadmap, and operating plan for AI model performance and inference quality. 
  • Establish performance budgets for latency, throughput, memory, GPU utilization, power, and numerical parity. 
  • Lead investigations into performance bottlenecks across model architecture, operators, kernels, memory movement, scheduling, runtime behavior, and hardware utilization. 
  • Establish repeatable benchmarking and profiling practices across simulation, hardware-in-the-loop, bench, and vehicle environments. 
  • Guide optimization through model architecture changes, operator and kernel improvements, memory optimization, scheduling, and hardware-aware execution. 
  • Build performance dashboards, regression detection, benchmark automation, and root-cause diagnostics. 
  • Partner with Embodied AI, model development, GPU kernel, runtime, system performance, vehicle integration, simulation, and safety teams. 
  • Influence model design by translating profiling results into clear recommendations for model architects and researchers. 
  • Represent AI Deployment in architecture reviews, program planning, and senior leadership discussions. 

Leadership Responsibilities 

  • Build and lead an inclusive, high-performing organization through hiring, coaching, feedback, and manager development. 
  • Establish clear ownership, priorities, staffing plans, and operating rhythms across performance workstreams. 
  • Define and manage KPIs for inference latency, latency variability, throughput, memory efficiency, GPU utilization, parity, and regression rate. 
  • Balance near-term production needs with longer-term investments in profiling, optimization automation, reduced precision, and performance infrastructure. 
  • Resolve cross-functional issues and align stakeholders when performance, quality, or implementation trade-offs are contested. 
  • Develop technical leaders and succession plans in GPU performance, model optimization, inference systems, and numerical analysis. 

Your Skills & Abilities (Required Qualifications) 

  • Bachelor’s degree in Computer Science, Electrical or Computer Engineering, Robotics, Machine Learning, or a related field; advanced degree preferred, or equivalent experience. 
  • 10+ years of experience in machine learning systems, model optimization, inference, GPU systems, robotics, autonomous driving, or a related field. 
  • 5+ years of people-leadership experience, including experience leading managers or senior technical leaders. 
  • Experience shipping production machine-learning inference systems on GPU, accelerator, robotics, automotive, or other edge hardware. 
  • Strong understanding of the factors that determine model performance: architecture, tensor shapes, operators, kernels, memory movement, scheduling, runtime execution, and hardware utilization. 
  • Hands-on experience with several of the following: PyTorch, CUDA, C++, Python, TensorRT, GPU profiling, benchmarking, performance analysis, or inference runtimes. 
  • Experience with quantization, pruning, distillation, architecture optimization, kernel optimization, or memory optimization. 
  • Experience building benchmark automation, performance regression detection, telemetry, dashboards, or profiling workflows. 
  • Strong systems thinking, communication, decision-making, and cross-functional leadership skills. 

What Will Give You a Competitive Edge 

  • Experience optimizing real-time machine-learning systems for autonomous driving, robotics, embedded systems, or computer vision. 
  • Deep experience with GPU performance, memory bandwidth, occupancy, synchronization, stream scheduling, or device-to-device data movement. 
  • Experience with NVIDIA Nsight Systems, NVIDIA Nsight Compute, PyTorch Profiler, TensorRT profiling tools, or equivalent tools. 
  • Experience deploying reduced-precision models and managing calibration, sensitivity, parity, and model-quality risks. 
  • Experience optimizing transformer, vision, lidar, or multimodal workloads. 
  • Experience measuring performance across simulation, hardware-in-the-loop, bench, and vehicle environments. 
  • Experience with safety-critical or highly reliable systems. 

Compensation: The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington

  • Compensation: The expected base compensation for this role is : $296,300 - $453,900 Actual base compensation within the identified range will vary based on factors relevant to the position.
  • Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance.
  • Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays

#GM-AV-1

이 직무는 재택 직무로 분류됩니다. 즉, 선발된 지원자는 근무 국가 어디에서든 근무할 수 있으며, 관리자의 지시가 없는 한 GM 사업장에 출근하지 않아도 됩니다.

선발된 지원자는 이 직무를 위해 25% 미만의 출장을 다녀야 합니다.

이 직무는 리로케이션 혜택을 받을 수 있습니다.

다양성 정보

General Motors는 법적으로 금지된 차별을 배제하는 것은 물론 포용성과 소속감을 진정으로 장려하는 직장이 되기 위해 노력하고 있습니다. 당사는 다양성이 보장되는 환경에서 직원들이 역량을 발휘하고 우리 고객을 위한 더 좋은 제품을 개발할 수 있다고 믿습니다. 따라서 입사에 관심 있는 사람이 있다면 포지션별 주요 업무와 자격을 확인하고 본인이 보유한 기술과 능력에 부합하는 모든 포지션에 적극적으로 지원하기를 장려합니다. 지원자는 채용 과정에서 역할 관련 평가(해당하는 경우) 및/또는 채용 전 스크리닝을 통과해야 합니다.  자세한 정보는 GM 채용 과정 안내를 참고하십시오.

공평한 취업 기회 선언 (미국)

General Motors는 공평한 기회를 제공하는 고용주임을 자부합니다.  자격을 만족하는 지원자는 인종과 피부색, 성별, 성적 지향, 성별 정체성, 국적, 장애, 재향 군인 보호법 적용 여부와 상관없이 채용 후보로서 심사를 받습니다. 

숙소 (미국 및 캐나다)

General Motors는 장애인을 포함한 모든 구직자들에게 취업 기회를 제공합니다. 구직이나 취업 지원에 도움이 되는 합리적인 숙소가 필요한 경우 [email protected]으로 이메일을 보내시거나 800-865-7580으로 전화주십시오. 이메일에, 귀하가 요청하는 특정한 숙소에 대한 설명과 귀하가 지원하는 직무와 채용 요청서 번호를 포함해주세요.