Search for More Jobs
Get alerts for jobs like this Get jobs like this tweeted to you
Company: AMD
Location: Austin, TX
Career Level: Mid-Senior Level
Industries: Technology, Software, IT, Electronics

Description



ADVANCE YOUR CAREER. ADVANCE THE WORLD. 

At AMD, we believe technology has the power to solve the world's most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future. 

 

Whether you're designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we'll advance your career.



THE ROLE: 

As a Fellow, you will be accountable for defining and driving the end-to-end software optimization strategy to achieve industry-leading performance for our top-tier customers. You will play a pivotal role in GPU architecture feature design, usage, API design, optimization, and deployment,

 

THE PERSON: 

We are looking for a talented and motivated GPU kernel developer with broad and deep experience in compute and machine learning software development to deliver out of the box high performance ROCm core software to our customers. A great candidate will be a strong collaborator who is open to continuous learning and is excited about innovative solutions for ML and HPC on GPUs.

 

KEY RESPONSIBILITIES: 

  • Perform external benchmarking and research to define opportunities to improve GPU kernel performance on AMD machine learning products
  • Analyze and optimize the performance of GPU kernels, providing insight on optimization paths across the AI software stack.
  • Partner with top customers and hyperscalers to understand their unique workload requirements and deliver tailored architectural wins and software optimizations.
  • Collaborate across hardware architecture, compiler, and framework teams to influence future silicon features based on evolving AI workload trends.
  • Communicate and collaborate with key technical experts across AMD and with our partners and customers to improve ROCm applications, libraries, and tools, as well as hardware

PREFERRED EXPERIENCE: 

  • Strong understanding of modern model architectures (Transformer, Attention, KV Cache) and optimization techniques like quantization, speculative decoding, and FlashAttention.
  • Excellent GPU architecture knowledge, performance analysis, algorithm development, and machine learning experience
  • Proven history of optimizing GPU kernels with advanced algorithms and across entire AI software stack.
  • Demonstrated ability to drive cross-functional initiatives in fast-paced, ambiguous environments.

PREFERRED ACADEMIC CREDENTIALS: 

  • Bachelor's or Master's or PhD degree in Computer Science, Computer Engineering, or equivalent.

LOCATION: Austin, Texas 

 

#LI-HYBRID

 

This role is not eligible for visa sponsorship.



Benefits offered are described:  AMD benefits at a glance.

 

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.

 

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD's “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.


 Apply on company website