Descrição
About the Organization
General Motors is developing the software and artificial intelligence capabilities for the next generation of autonomous driving. Within AI Foundations, AI Acceleration makes machine learning models faster, more efficient, and more reliable on production vehicle hardware.
The AI Deployment team owns model inference performance across simulation, hardware-in-the-loop, bench, and vehicle environments. The team focuses on latency, memory, GPU utilization, numerical parity, profiling, benchmarking, reduced precision, and production readiness.
About the Role
We are looking for a Senior Manager, AI Deployment to lead the strategy and execution of model performance and on-vehicle inference for autonomous driving. You will lead engineering managers and senior technical leaders working across model optimization, GPU systems, inference runtimes, and vehicle integration. You will set performance goals, guide optimization of complex autonomy models, and establish disciplined methods to measure latency, diagnose regressions, and validate improvements. Success requires strong technical judgment, people leadership, and the ability to make clear trade-offs among latency, memory, throughput, accuracy, power, and numerical parity.
What You’ll Do
- Own the strategy, roadmap, and operating plan for AI model performance and inference quality.
- Establish performance budgets for latency, throughput, memory, GPU utilization, power, and numerical parity.
- Lead investigations into performance bottlenecks across model architecture, operators, kernels, memory movement, scheduling, runtime behavior, and hardware utilization.
- Establish repeatable benchmarking and profiling practices across simulation, hardware-in-the-loop, bench, and vehicle environments.
- Guide optimization through model architecture changes, operator and kernel improvements, memory optimization, scheduling, and hardware-aware execution.
- Build performance dashboards, regression detection, benchmark automation, and root-cause diagnostics.
- Partner with Embodied AI, model development, GPU kernel, runtime, system performance, vehicle integration, simulation, and safety teams.
- Influence model design by translating profiling results into clear recommendations for model architects and researchers.
- Represent AI Deployment in architecture reviews, program planning, and senior leadership discussions.
Leadership Responsibilities
- Build and lead an inclusive, high-performing organization through hiring, coaching, feedback, and manager development.
- Establish clear ownership, priorities, staffing plans, and operating rhythms across performance workstreams.
- Define and manage KPIs for inference latency, latency variability, throughput, memory efficiency, GPU utilization, parity, and regression rate.
- Balance near-term production needs with longer-term investments in profiling, optimization automation, reduced precision, and performance infrastructure.
- Resolve cross-functional issues and align stakeholders when performance, quality, or implementation trade-offs are contested.
- Develop technical leaders and succession plans in GPU performance, model optimization, inference systems, and numerical analysis.
Your Skills & Abilities (Required Qualifications)
- Bachelor’s degree in Computer Science, Electrical or Computer Engineering, Robotics, Machine Learning, or a related field; advanced degree preferred, or equivalent experience.
- 10+ years of experience in machine learning systems, model optimization, inference, GPU systems, robotics, autonomous driving, or a related field.
- 5+ years of people-leadership experience, including experience leading managers or senior technical leaders.
- Experience shipping production machine-learning inference systems on GPU, accelerator, robotics, automotive, or other edge hardware.
- Strong understanding of the factors that determine model performance: architecture, tensor shapes, operators, kernels, memory movement, scheduling, runtime execution, and hardware utilization.
- Hands-on experience with several of the following: PyTorch, CUDA, C++, Python, TensorRT, GPU profiling, benchmarking, performance analysis, or inference runtimes.
- Experience with quantization, pruning, distillation, architecture optimization, kernel optimization, or memory optimization.
- Experience building benchmark automation, performance regression detection, telemetry, dashboards, or profiling workflows.
- Strong systems thinking, communication, decision-making, and cross-functional leadership skills.
What Will Give You a Competitive Edge
- Experience optimizing real-time machine-learning systems for autonomous driving, robotics, embedded systems, or computer vision.
- Deep experience with GPU performance, memory bandwidth, occupancy, synchronization, stream scheduling, or device-to-device data movement.
- Experience with NVIDIA Nsight Systems, NVIDIA Nsight Compute, PyTorch Profiler, TensorRT profiling tools, or equivalent tools.
- Experience deploying reduced-precision models and managing calibration, sensitivity, parity, and model-quality risks.
- Experience optimizing transformer, vision, lidar, or multimodal workloads.
- Experience measuring performance across simulation, hardware-in-the-loop, bench, and vehicle environments.
- Experience with safety-critical or highly reliable systems.
Compensation: The compensation information is a good faith estimate only. It is based on what a successful applicant might be paid in accordance with applicable state laws. The compensation may not be representative for positions located outside of New York, Colorado, California, or Washington
- Compensation: The expected base compensation for this role is : $296,300 - $453,900 Actual base compensation within the identified range will vary based on factors relevant to the position.
- Bonus Potential: An incentive pay program offers payouts based on company performance, job level, and individual performance.
- Benefits: GM offers a variety of health and wellbeing benefit programs. Benefit options include medical, dental, vision, Health Savings Account, Flexible Spending Accounts, retirement savings plan, sickness and accident benefits, life insurance, paid vacation & holidays
#GM-AV-1
Esta função é classificada como remota. Isso significa que o candidato selecionado pode estar baseado em qualquer lugar do país de trabalho e não deve se apresentar em um local de trabalho da GM, a menos que seja orientado por seu gerente.
O candidato selecionado deverá viajar <25% para esta função.
Esta posição pode ser elegível para benefícios de relocação.
Informações sobre diversidade
A General Motors está comprometida em ser um local de trabalho que não só é livre de discriminação ilegal, como estimula verdadeiramente a inclusão e integração. Acreditamos enfaticamente que a diversidade na força de trabalho cria um ambiente no qual nossos colaboradores podem crescer e desenvolver melhores produtos para nossos clientes. Incentivamos os candidatos interessados a analisar as principais responsabilidades e qualificações de cada função e a se candidatar a qualquer cargo que corresponda a suas habilidades e capacidades. Os candidatos no processo de recrutamento podem, quando aplicável, ser solicitados a concluir com sucesso uma ou mais avaliações relacionadas à função e/ou uma seleção pré-emprego antes de iniciar o emprego. Para saber mais, acesse Como contratamos.
Declaração de Igualdade de Oportunidades de Emprego (EUA)
A General Motors tem orgulho de ser um empregador que oferece oportunidades iguais. Todos os candidatos qualificados serão considerados para o emprego, independentemente de raça, cor, religião, sexo, orientação sexual, identidade de gênero, origem nacional, deficiência ou status como veterano protegido.
Adaptações (EUA e Canadá)
A General Motors oferece oportunidades a todos os candidatos a emprego, incluindo pessoas com deficiências. Se você precisa de uma adaptação razoável para ajudá-lo na sua pesquisa de cargos ou solicitação de emprego, fale conosco pelo e-mail [email protected] ou pelo telefone 800-865-7580. No seu e-mail, inclua uma descrição da adaptação específica que você está solicitando assim como o nome do cargo e o número de requisição do cargo ao qual está se candidatando.
