Talent.com
Advanced Micro Devices, Inc
Staff Software Development Engineer: GPU, Computer Vision, AI/ML OpsAdvanced Micro Devices, Inc • Santa Clara, California, United States
Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops

Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops

Advanced Micro Devices, Inc • Santa Clara, California, United States
25 days ago
Job type
  • Full-time
Job description


WHAT YOU DO AT AMD CHANGES EVERYTHING

At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond.Together, we advance your career.




THE ROLE:

AMD is looking for an influential software engineer who is passionate about improving the performance of key applications and benchmarks. You will be a member of a core team of incredibly talented industry specialists and will work with the very latest hardware and software technology.

THE PERSON:

As a Staff Software Developer, you will be at the heart of AMD's AI strategy, tackling one of the most exciting challenges in the industry: training and running AI to make AI itself more efficient on GPUs on the fly, which can dramatically alter the trajectory of AI progress. This is a high-impact, hands-on role where your work will directly define the software that powers the future of AI.

KEY RESPONSIBILITIES:

Architect and Drive the AI Software Stack: You will establish best practices and optimize performance from the lowest-level GPU kernels to large-scale distributed systems, shaping the foundational software for AMD hardware. By leveraging cutting-edge Large Language Models (LLMs) and agent-based technologies, you will accelerate the development and performance enhancement of the AMD ROCm ecosystem, ensuring it remains at the forefront of AI innovation.


Accelerate Foundational Models: Your work will directly accelerate cutting-edge applications like foundation models (LLMs) and autonomous AI agents, ensuring AMD is the platform of choice for the most demanding workloads.


Innovate Across Hardware and Software: You will contribute to the entire co-design lifecycle, from influencing future GPU architectures to developing groundbreaking software for new accelerators and collaborating with the broader AI community.


Success in this role requires a deep passion for software engineering, strong technical ownership to see complex problems through to resolution, and the ability to influence technical direction across teams. As a senior engineer, you will also be expected to mentor others and effectively communicate your ideas to shape the future of AI at AMD.


To excel in this role, we seek a candidate with exceptional technical expertise, who can bridge deep proficiency in high-performance C++ software engineering and low-level GPU programming with a robust understanding of Large Language Models (LLMs) and AI systems. The ideal candidate can bridge kernel engineering with AI post-training (RL) experience. A great candidate is deep in one and light on the other.


Kernel engineering means demonstrating mastery in designing complex, scalable systems using modern C++, coupled with a fundamental grasp of GPU architectures (HIP/CUDA), memory hierarchies, and kernel optimization to maximize hardware performance. This expertise should be evidenced by significant hands-on experience in large-scale C++/HIP/CUDA projects, such as contributing to the ROCm ecosystem (e.g., rpp, MIVisionX, rocAL, rocdecode, rocjpeg), CUDA libraries (e.g., CV-CUDA, cuDNN, NCCL), or the C++/HIP/CUDA core of ML frameworks like PyTorch, TensorFlow, or JAX.


AI post-training is equally critical, and requires deep understanding of LLMs, including but not limited to transformer architectures, attention mechanisms, and the full model lifecycle, with hands-on experience in advanced model alignment and post-training techniques like Supervised Fine-Tuning (SFT) and Reinforcement Learning (e.g., RLHF, GRPO). Candidates must also stay at the forefront of LLM advancements, showing familiarity with cutting-edge trends such as Mixture-of-Experts (MoE) architectures, inference optimizations (e.g., quantization, speculative decoding), and modern application patterns like Agentic AI systems (e.g. AlphaEvolve for code/kernel generation).


Experience and interest in code generation and/or self-improving LLMs is a plus.

PREFERRED EXPERIENCE:

  • This is a senior role that requires a unique blend of expertise across software engineering, GPU computing, and artificial intelligence. The ideal candidate will possess:

    Lengthy professional software development experience in performance-critical environments.
    Extensive hands-on experience in GPU programming (HIP/CUDA) and optimizing deep learning kernels and operators.Computer vision expertiseA fundamental understanding of GPU architecture and memory hierarchy, used to diagnose and resolve complex performance bottlenecks.Expert-level proficiency in modern C++ and object-oriented design.Deep experience using GPU profiling and performance analysis tools (e.g., AMD ROCm Profiler, NVIDIA Nsight) to diagnose and resolve complex bottlenecks in distributed, multi-GPU systems.Deep knowledge of transformer architectures, attention mechanisms, and modern AI systems (Generative AI, Agentic AI).Hands-on experience optimizing the post-training and inference pipelines of Large Language Models (LLMs).Strong technical ownership, communication, and problem-solving skills with a track record of delivering complex technical solutions.Plus: Experience or deep expertise with the AMD ROCm/HIP ecosystem.

ACADEMIC CREDENTIALS:

  • Bachelor’s or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent
  • Relevant publications in AI/ML, GPU computing, or system optimization are highly valued.

This role is not eligible for visa sponsorship.




Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

THE ROLE:

AMD is looking for an influential software engineer who is passionate about improving the performance of key applications and benchmarks. You will be a member of a core team of incredibly talented industry specialists and will work with the very latest hardware and software technology.

THE PERSON:

As a Staff Software Developer, you will be at the heart of AMD's AI strategy, tackling one of the most exciting challenges in the industry: training and running AI to make AI itself more efficient on GPUs on the fly, which can dramatically alter the trajectory of AI progress. This is a high-impact, hands-on role where your work will directly define the software that powers the future of AI.

KEY RESPONSIBILITIES:

Architect and Drive the AI Software Stack: You will establish best practices and optimize performance from the lowest-level GPU kernels to large-scale distributed systems, shaping the foundational software for AMD hardware. By leveraging cutting-edge Large Language Models (LLMs) and agent-based technologies, you will accelerate the development and performance enhancement of the AMD ROCm ecosystem, ensuring it remains at the forefront of AI innovation.


Accelerate Foundational Models: Your work will directly accelerate cutting-edge applications like foundation models (LLMs) and autonomous AI agents, ensuring AMD is the platform of choice for the most demanding workloads.


Innovate Across Hardware and Software: You will contribute to the entire co-design lifecycle, from influencing future GPU architectures to developing groundbreaking software for new accelerators and collaborating with the broader AI community.


Success in this role requires a deep passion for software engineering, strong technical ownership to see complex problems through to resolution, and the ability to influence technical direction across teams. As a senior engineer, you will also be expected to mentor others and effectively communicate your ideas to shape the future of AI at AMD.


To excel in this role, we seek a candidate with exceptional technical expertise, who can bridge deep proficiency in high-performance C++ software engineering and low-level GPU programming with a robust understanding of Large Language Models (LLMs) and AI systems. The ideal candidate can bridge kernel engineering with AI post-training (RL) experience. A great candidate is deep in one and light on the other.


Kernel engineering means demonstrating mastery in designing complex, scalable systems using modern C++, coupled with a fundamental grasp of GPU architectures (HIP/CUDA), memory hierarchies, and kernel optimization to maximize hardware performance. This expertise should be evidenced by significant hands-on experience in large-scale C++/HIP/CUDA projects, such as contributing to the ROCm ecosystem (e.g., rpp, MIVisionX, rocAL, rocdecode, rocjpeg), CUDA libraries (e.g., CV-CUDA, cuDNN, NCCL), or the C++/HIP/CUDA core of ML frameworks like PyTorch, TensorFlow, or JAX.


AI post-training is equally critical, and requires deep understanding of LLMs, including but not limited to transformer architectures, attention mechanisms, and the full model lifecycle, with hands-on experience in advanced model alignment and post-training techniques like Supervised Fine-Tuning (SFT) and Reinforcement Learning (e.g., RLHF, GRPO). Candidates must also stay at the forefront of LLM advancements, showing familiarity with cutting-edge trends such as Mixture-of-Experts (MoE) architectures, inference optimizations (e.g., quantization, speculative decoding), and modern application patterns like Agentic AI systems (e.g. AlphaEvolve for code/kernel generation).


Experience and interest in code generation and/or self-improving LLMs is a plus.

PREFERRED EXPERIENCE:

  • This is a senior role that requires a unique blend of expertise across software engineering, GPU computing, and artificial intelligence. The ideal candidate will possess:

    Lengthy professional software development experience in performance-critical environments.
    Extensive hands-on experience in GPU programming (HIP/CUDA) and optimizing deep learning kernels and operators.Computer vision expertiseA fundamental understanding of GPU architecture and memory hierarchy, used to diagnose and resolve complex performance bottlenecks.Expert-level proficiency in modern C++ and object-oriented design.Deep experience using GPU profiling and performance analysis tools (e.g., AMD ROCm Profiler, NVIDIA Nsight) to diagnose and resolve complex bottlenecks in distributed, multi-GPU systems.Deep knowledge of transformer architectures, attention mechanisms, and modern AI systems (Generative AI, Agentic AI).Hands-on experience optimizing the post-training and inference pipelines of Large Language Models (LLMs).Strong technical ownership, communication, and problem-solving skills with a track record of delivering complex technical solutions.Plus: Experience or deep expertise with the AMD ROCm/HIP ecosystem.

ACADEMIC CREDENTIALS:

  • Bachelor’s or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent
  • Relevant publications in AI/ML, GPU computing, or system optimization are highly valued.

This role is not eligible for visa sponsorship.

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Create a job alert for this search

Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops • Santa Clara, California, United States

Similar jobs

Senior Staff Software Engineer, Infotainment (Android Applications & Frameworks)

Scout MotorsFremont, CA, United States
Full-time

Senior Staff Software Engineer, Infotainment (Android Applications & Frameworks)Here at Scout Motors, we're carrying forward the heritage of one of the most iconic American vehicles in history.... Show more

 • Promoted

Member of Technical Staff 2- AI/ML

NutanixSan Jose, CA, United States
Full-time

Hungry, Humble, Honest, with Heart.The OpportunityAre you an AI/ML engineer passionate about building intelligent systems from the ground up? Join the SaaS Engineering team at Nutanix to design, de... Show more

 • Promoted

Software Engineer Backend Development:

AkrayaFremont, CA, United States
Full-time

Software Engineer OpportunityPrimary Skills:JavaScript (Expert), Go (Advanced), Python (Intermediate), ASP.NET (Expert), REST APIs (Advanced).Contract Type:W2/C2C Duration:6Months Location:Fremont,... Show more

 • Promoted

Field Sales Veterinary Diagnostics Santa Cruz CA

IDEXX LaboratoriesSanta Cruz, CA, United States
Full-time

Communicating the true value of our veterinary diagnostic and technology products and services is at the heart of our IDEXX's commercial business.Our sales professionals develop deep and meaningful... Show more

 • Promoted

Staff Software Engineer - Remote

TradeJobsWorkForce95122 San Jose, CA, US
Remote
Full-time

Staff Software Engineer Remote Job Duties: • Implement and evolve a Data Lake storage system with low latency and high thr... Show more

 • Promoted

Medical Billing and Coding - Entry Level Training Program

Dreambound Inc.Capitola, California, US
Full-time

Note : This is an educational program, not a job.Successful completion of the program does not guarantee employment but will equip you with valuable skills for the healthcare job market.Looking to ... Show more

 • Promoted

Staff Engineering Program Manager

Heron PowerScotts Valley, CA, United States
Full-time

Staff Engineering Program Manager.Heron Power is a startup company building cutting-edge power electronics for the 21st-century grid.We aim to debottleneck the growth of electricity generation and ... Show more

 • Promoted

Member of Technical Staff - Software Engineer (SuperIntelligence team)

Microsoft CorporationMountain View, CA, United States
Full-time

OverviewHelp build the infrastructure that powers training, evaluation, and data platforms for reliable deployment of world-class foundational AI models.We are on a mission to create state-of-the-a... Show more

 • Promoted

Software Engineer, Machine Learning

Meta PlatformsNewark, CA, United States
Full-time

Talented Engineers WantedMeta is seeking talented engineers to join our teams in building cutting-edge products that connect billions of people around the world.As a member of our team, you will ha... Show more

 • Promoted

Senior Software Engineer, Flight Simulator

Joby AviationSanta Cruz, CA, United States
Full-time

Joby Flight Training Simulator EngineerImagine a piloted air taxi that takes off vertically, then quietly carries you and your fellow passengers over the congested city streets below, enabling you ... Show more

 • Promoted

Staff Software Engineer, Data Frameworks

Ridge Line ServicesSan Ramon, CA, United States
Full-time

Staff Software Engineer, Data FrameworksAre you a seasoned engineer passionate about transforming data into a powerful business asset? Do you thrive in environments where you can lead by example, m... Show more

 • Promoted

Senior Manager, Gen AI Software Engineering

Thermo FisherPleasanton, CA, United States
Full-time

Senior Manager, Generative AI Software EngineeringAs part of the Thermo Fisher Scientific team, you'll discover meaningful work that makes a positive impact on a global scale.Join our colleagues in... Show more

 • Promoted

Travel Radiology Tech - $2,808 per week in Santa Cruz, CA

AlliedTravelCareersSanta Cruz, CA, US
Full-time

AlliedTravelCareers is working with AHS Staffing to find a qualified Rad Tech in Santa Cruz, California, 95065!.AHS Staffing is looking for a CT Tech Radiologic Technologist in Santa Cruz, CA for a... Show more

 • Promoted

Staff Software Engineer - Key Management and Cryptography

LinkedInSunnyvale, CA, United States
Full-time

Staff Software Engineer - Key Management and CryptographyLinkedIn is the world's largest professional network, built to create economic opportunity for every member of the global workforce.Our prod... Show more

 • Promoted

Staff Software Engineer

AbbottPleasanton, CA, United States
Full-time

Staff Software EngineerAbbott is a global healthcare leader that helps people live more fully at all stages of life.Our portfolio of life-changing technologies spans the spectrum of healthcare, wit... Show more

 • Promoted

Physician (MD/DO) - Pediatrics - General/Other - $189,000 to $272,000 per year in Santa Cruz, CA

LocumJobsOnlineSanta Cruz, CA, US
$32.00 hourly
Full-time +2

Doctor of Medicine | Pediatrics - General/Other.LocumJobsOnline is working with CompHealth to find a qualified Pediatrics MD in Santa Cruz, California, 95062!.Santa Cruz, CA offers physicians the r... Show more

 • Promoted

Staff AI Research Engineer

Advanced Micro Devices, Inc.Santa Clara, CA, United States
Full-time

What You Do At AMD Changes EverythingAt AMD, our mission is to build great products that accelerate next-generation computing experiencesfrom AI and data centers, to PCs, gaming and embedded system... Show more

 • Promoted

Senior Software Engineer (full-stack)

Charge RoboticsSan Leandro, CA, United States
Full-time

Charge RoboticsCharge Robotics is a Series A startup building robots that build solar farms.Demand for new solar projects is booming (1 ? 5 of all the solar that exists in the US was installed last... Show more

 • Promoted

AGM Santa Cruz: $80k - $90k

Foley and FitzgeraldSanta Cruz, CA, United States
Full-time

This position requires fine-dining, high-end service, FOH leadership experience.The restaurant is on track to one Michelin star, and the service execution needs to match the food execution.Join an ... Show more

 • Promoted

Training and Development Associate - Santa Cruz

GOODWILL CENTRAL COASTSanta Cruz, CA, United States
Full-time

Training And Development Associate - Santa CruzSalary Range:$17.HourlyPosition Type:Full TimeJob Shift:DayEducation Level:High SchoolDescriptionJob Summary:The Training & Development Associate ... Show more