High Performance and Scientific Computing Fellow

Advanced Micro Devices, IncAustin, TX
Hybrid

About The Position

The AI GPU Software (AGS) team is looking for a High Performance and Scientific Computing Fellow. This is a technical leadership position as an Individual Contributor (IC) driving innovation across multiple teams developing High Performance Computing (HPC) GPU libraries as part of the AMD ROCm™ Software platform. You will focus primarily on deep technical leadership rather than organizational management. As a Fellow, you will be a primary contributor towards defining software strategy and driving the technical vision for HPC GPU libraries across multiple generations of AMD GPUs. This role will require deep expertise in mathematical algorithms, floating point precision, GPU performance analysis, distributed systems, software architecture and engineering. The ideal candidate is highly hands-on and embraces agentic AI workflows and is expected to play a key role in industry-standard benchmarks and external technical engagements.

Requirements

  • Accustomed to working in a dynamic, geographically distributed agile team, where partnership and collaboration are paramount.
  • Excellent written and verbal communication skills, strong attention to detail, and the ability to express your work in a clear, cohesive fashion.
  • Recognized technical leader with contributions across the HPC software stack and applications, both at the node and distributed level.
  • Understand how to architect optimal software solutions for HPC customers while balancing often conflicting priorities.
  • Comfortable operating across layers—from kernels and runtimes to libraries and distributed strategies—and have a track record of driving impactful optimizations and influencing technical direction.
  • In depth understanding of mathematical algorithms for dense and sparse linear algebra.
  • Experience with direct and iterative solvers.
  • Experience designing, implementing, debugging, and optimizing parallel algorithms on class-leading supercomputers.
  • Knowledge of important HPC algorithms and libraries.
  • Strong background developing applications and libraries in C++, C and Fortran.
  • Familiarity with GPU software development and optimization using HIP, CUDA, or OpenCL and distributed programming with MPI and/or SHMEM.
  • Understanding of CPU and GPU architectures and low-level optimization techniques including assembly programming and vectorization.
  • In-depth knowledge of best-practices in software development, including testing, profiling, debugging, documentation, version control, issue tracking, and planning.

Nice To Haves

  • Advanced degrees, such as M.Sc., M.Eng., Ph.D. are preferred.

Responsibilities

  • Define technical vision and drive software strategy across AMD ROCm HPC libraries.
  • Lead performance analysis, tuning, and algorithmic innovation across HPC libraries.
  • Provide technical mentorship to senior engineers and influence best practices across the organization.
  • Communicate complex technical findings and recommendations to senior leadership and stakeholders.
  • Represent AMD in external technical forums, benchmarks, and customer engagements.
  • Work with key technical experts across AMD and with our partners and customers to improve ROCm HPC and AI applications, libraries, and tools.

Benefits

  • AMD benefits at a glance.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service