ML GPU Kernel Development Engineer
2 weeks ago
WHAT YOU DO AT AMD CHANGES EVERYTHING
At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you'll discover the real differentiator is our culture. We push the limits of innovation to solve the world's most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond.
Together, we advance your career.
ML GPU Kernel Development Engineer
The Role
We are seeking a talented Machine Learning Kernel Developer to design, develop, and optimize low-level machine learning kernels for AMD GPUs using the ROCm software stack. In this role, you will work on high-impact projects to accelerate AI frameworks and libraries, with a focus on emerging technologies like Large Language Models (LLMs) and other generative AI workloads.
The Person
The ideal candidate will have hands-on experience with GPU programming (ROCm or CUDA) and a passion for pushing the boundaries of AI performance.
Key Responsibilities
- Design and implement highly optimized ML kernels (e.g., matrix operations, attention mechanisms) for AMD GPUs using ROCm.
- Profile, debug, and tune kernel performance to maximize hardware utilization for AI workloads.
- Collaborate with ML researchers and framework developers to integrate kernels into AI frameworks (e.g., PyTorch, TensorFlow) and inference engines (e.g., vLLM, SGLang).
- Contribute to the ROCm software stack by identifying and resolving bottlenecks in libraries like MIOpen, BLAS, or Composable Kernel.
- Stay updated on the latest AI/ML trends (LLMs, quantization, distributed inference) and apply them to kernel development.
- Document and communicate technical designs, benchmarks, and best practices.
- Troubleshoot and resolve issues related to GPU compatibility, performance, and scalability.
Required Experience
- 2+ years of experience in GPU kernel development for machine learning (ROCm or CUDA).
- Proficiency in C/C++ and Python, with experience in performance-critical programming.
- Strong understanding of ML frameworks (PyTorch, TensorFlow) and GPU-accelerated libraries.
- Basic knowledge of modern AI technologies (LLMs, transformers, inference optimization).
- Familiarity with parallel computing, memory optimization, and hardware architectures.
- Problem-solving skills and ability to work in a fast-paced environment.
Preferred Experience
- Direct experience with AMD ROCm development (HIP, MIOpen, Composable Kernel).
- Knowledge of LLM-specific optimizations (e.g., FlashAttention, PagedAttention in vLLM).
- Experience with distributed training/inference or model compression techniques.
- Contributions to open-source ML projects or GPU compute libraries.
Academic Credentials
- Bachelor's/Master's in Computer Science, Electrical Engineering, or related field.
Benefits offered are described:
AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.
-
ML GPU Kernel Development Engineer
2 weeks ago
Hyderabad, Telangana, India AMD Full timeOverview:WHAT YOU DO AT AMD CHANGES EVERYTHINGAt AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to...
-
ML GPU Kernel Development Engineer
2 weeks ago
Hyderabad, Telangana, India Advanced Micro Devices, Inc Full timeWHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create...
-
Open source AI/ML
2 weeks ago
Hyderabad, Telangana, India Source-Right Full timePosition: Open source AI/ML (SI35FT RM 3718)EXPERIENCE – Must HaveStrong C++ and Python programming skills.Performance analysis skills for both CPU and GPUGood knowledge of AI/ML Frameworks and ArchitectureBasic GPU kernel programming knowledgeExperience with software engineering methodologies such as Agile, Scrum, Kanban.Experience in all the phases of...
-
GPU Developer/ Engineer
6 days ago
Hyderabad, Telangana, India Mirafra Full timeTitle: GPU Developers/ GPU validation Engineer/ LeadsLocation: Hyderabad or BangaloreDescription:C++ programmingExperience in GPU Architectures, GPU Pipelines, GPU game processing, GPU rendering image processingExperience in OpenCL, Open GL, Vulkan and profilingGFX testing, Sanity/Stability/regression and performance testing
-
Open source AI/ML
2 weeks ago
Hyderabad, Telangana, India Source-Right Inc. Full timePosition: Open source AI/ML (SI35FT RM 3718)EXPERIENCE – Must have :Strong C++ and Python programming skills.Performance analysis skills for both CPU and GPUGood knowledge of AI/ML Frameworks and ArchitectureBasic GPU kernel programming knowledgeExperience with software engineering methodologies such as Agile, Scrum, Kanban.Experience in all the phases of...
-
GPU Validation Engineer
2 weeks ago
Hyderabad, Telangana, India ElevarSoC Technologies Full timeHello EveryoneGreetings from ElevarSoCWe are hiring for GPU Validation Engineer for Hyderabad location with 1-3 Years of experienceBelow the JdGood knowledge on Graphics Device drivers WHQL validation, Python based Automation Execution, 3D gaming and multimediaExperience with Test Automation skills like Python, C and other scripting languagesKnowledge of...
-
AI/ML Engineer
2 days ago
Hyderabad, Telangana, India Infosif solution Full timeJDAI/ML ENgineerl Hands-on experience with NVIDIA GPU acceleration, CUDA, TensorRT, and deep learning frameworks (e.g., PyTorch, TensorFlow).]l Vision + deep learning + sensor fusionl ML Models, ML Infrastructurel ROSl Python, Linux, and modern development workflows (Git, CI/CD, etc.).Notice period immediate to 20 daysJob Type: Full-timeWork Location: In...
-
Security Kernel Developer
5 days ago
Hyderabad, Telangana, India Onzestt Services Full timeRole & responsibilitiesEmbedded C programming (hands-on)Linux device driver developmentKernel security experience (12 years)Strong understanding of Linux kernel architectureKernel-level debugging skillsClient interaction and stakeholder communicationAbility to work independently on assigned modulesDevelop and enhance kernel security featuresDesign and...
-
AI/ML Engineer
2 weeks ago
Hyderabad, Telangana, India Infosif solution Full timeJd for AI/ML Engineerl Hands-on experience with NVIDIA GPU acceleration, CUDA, TensorRT, and deep learning frameworks (e.g., PyTorch, TensorFlow).]l Vision + deep learning + sensor fusionl ML Models, ML Infrastructurel ROSl Python, Linux, and modern development workflows (Git, CI/CD, etc.).Exp 5 + years onlyLocation HyderabadNotice period-immediate to 20...
-
AI/ML Engineer
2 weeks ago
Hyderabad, Telangana, India HiringNinja Full timeNight shift starting 9.30 pm istResponsibilities● Develop, train, and optimize AI/ML models that enhance Flippy's perception,decision-making, and real-time operational performance.● Conduct experiments and research to improve model accuracy, robustness, and generalization across diverse kitchen environments.● Collaborate with software, hardware, and...