Talent.com
Mercor
Remote CUDA Engineering Expert - AI Trainer ($500-$500 per hour)Mercor • Fremont, California, US
Remote CUDA Engineering Expert - AI Trainer ($500-$500 per hour)

Remote CUDA Engineering Expert - AI Trainer ($500-$500 per hour)

Mercor • Fremont, California, US
9 days ago
Job type
  • Full-time
  • Remote
Job description
## **1\. Role Overview** Mercor is seeking GPU kernel optimization experts to contribute to a project with a leading AI lab. This opportunity is designed for freelancers with strong C++ skills, practical GPU programming experience, and the ability to improve kernel performance using profiler-guided analysis. You’ll help evaluate, optimize, and reason about GPU kernels across modern hardware environments. This is a contract-based opportunity for specialists who enjoy squeezing performance out of modern GPU architectures. ## **2\. Key Responsibilities** - Analyze and optimize GPU kernels for performance, efficiency, and hardware utilization - Use profiler metrics such as L2 cache hit rate, L2 throughput, occupancy, and related signals to guide kernel improvements - Review GPU kernel implementations and identify bottlenecks without requiring extensive background in the underlying algorithms - Write, modify, and reason about C++17, Python, and GPU programming code - Apply CUDA, HIP, shader programming, or related kernel programming expertise to improve performance outcomes - Document optimization decisions clearly, including when specific profiler metrics are or are not useful ## **3\. Ideal Qualifications** - Available to work at least 20 hrs/wk - Fluent in core C++ features through C++17 - Working knowledge of Python and Git - Fluent in at least one GPU programming model, such as CUDA, HIP, Slang, HLSL, GLSL, or related kernel programming - At least 1 year of professional or graduate-level research experience working with GPUs - Strong understanding of GPU profiler performance metrics and how to use them to optimize kernels - Ability to optimize GPU kernels without needing deep prior context on every algorithm - Experience with CUDA, HIP, CUDA C++ Core Libraries, inline PTX assembly, or tensor core-level optimization is a plus - Experience optimizing kernels for NVIDIA Blackwell hardware is a plus - Familiarity with NSight Compute is a plus - Prior experience with GPU hardware organizations such as NVIDIA, AMD, or Qualcomm is a plus - Open-source contributions related to GPU kernel optimization are a plus ## **4\. Application Process** - Submit your resume or relevant technical background to get started - Qualified applicants may be asked to complete a brief technical assessment or submit additional information
Create a job alert for this search

Remote CUDA Engineering Expert - AI Trainer ($500-$500 per hour) • Fremont, California, US

Similar jobs

Senior Fullstack Engineer -- AI-Driven Healthcare (Remote)

MachinifyPalo Alto, CA, United States
Remote
Full-time

A leading healthcare intelligence company in Palo Alto is seeking a Staff or Sr.Fullstack Engineer to join their engineering team.You will design and build scalable web applications, leveraging you... Show more

 • Promoted

Remote Physics Expert (PhD) - $100+ per hour

Turing Benicia, California
Remote
Full-time
Quick Apply

Remote contract for PhDs in Physics, Applied Physics, or related fields.Work on cutting-edge projects with top AI labs while earning $100+/hour, fully remote, with flexible weekly hours.Help fine-t... Show more

 • Promoted

Remote Mathematics Expert (PhD) - $100+ per hour

Turing Pittsburg, California
Remote
Full-time
Quick Apply

Remote contract for PhDs in Mathematics, Statistics, or related fields.Work on cutting-edge projects with top AI labs while earning $100+/hour, fully remote, with flexible weekly hours.Help fine-tu... Show more

 • Promoted

DevOps Engineer AI Trainer

DataAnnotationPalo Alto, California, United States
$104,000.00 yearly
Full-time

Salary: $104,000 - 208,000 per year.We require native or bilingual fluency in English.We require proficiency in at least one of the following programming languages or frameworks: JavaScript, TypeSc... Show more