Search for More Jobs
Get alerts for jobs like this Get jobs like this tweeted to you
Company: AMD
Location: HKI, Uusimaa, Finland
Career Level: Mid-Senior Level
Industries: Technology, Software, IT, Electronics

Description



ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believe technology can change lives for the better. It can heal us, entertain us, and make us more connected, productive, and understanding of the world around us. And we're looking for talent who feel the same: people who want to leave the planet better than they found it, those who don't shy away from humanity's challenges but are determined to help solve them.

 

AMD is powering the next generation of supercomputing, high-performance computing, cloud, and AI. Whether you're designing next-gen processors, enabling AI breakthroughs, or creating go-to-market plans, every role at AMD contributes to something bigger — technology that moves the world forward.



Senior Agentic System & Application Engineer Role Summary

We are looking for an experienced engineer to design and develop agentic AI systems that automate complex HPC and enterprise workflows. In this role, you will build production-grade agents capable of planning, invoking tools, managing workflow state, launching jobs, diagnosing failures, optimizing execution, and operating safely across shared compute and enterprise environments.

The successful candidate will have experience delivering reliable LLM-powered agentic systems, a strong understanding of infrastructure and developer workflows, and the ability to bridge agentic AI, distributed systems, HPC scheduling, enterprise automation, observability, and performance optimization.

Key Responsibilities
  • Design agents for HPC and enterprise workflows, incorporating planning, tool use, memory, retrieval, governance, and human approval mechanisms.
  • Build production systems for job orchestration, enterprise workflow automation, monitoring, debugging, optimization, reproducibility, and rollback.
  • Automate environments, containers, toolchains, schedulers, CI/CD, enterprise systems, telemetry, workflow composition, and launch processes.
  • Implement safe execution through authentication, authorization, secrets handling, audit logging, sandboxing, compliance controls, and policy enforcement.
  • Research and prototype agentic methods for long-running workflows, failure recovery, optimization loops, evaluation, and human-in-the-loop reliability.
Collaboration
  • Partner with Platform teams to deploy solutions safely on shared clusters, enterprise infrastructure, and governed execution environments.
  • Work with Performance teams to define benchmarks, variance controls, profiling methods, productivity metrics, and acceptance criteria.
  • Collaborate with Compiler and Runtime teams to expose code generation, tuning, tracing, debugging, and execution controls.
  • Engage workload owners, enterprise application teams, security, and compliance stakeholders to onboard use cases and ensure auditable operation.
What You'll Bring
  • Advanced degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field, or equivalent practical experience.
  • Experience delivering production agentic AI or LLM systems with orchestration, tool use, memory, evaluations, and long-horizon reliability.
  • Strong proficiency in Python and experience with at least one systems programming language such as C, C++, Rust, or Go.
  • Experience integrating automation with real-world tooling, including code execution, build/test systems, schedulers, enterprise systems, telemetry, CI/CD, or deployment pipelines.
  • Experience building reliable production systems with observability, fallback mechanisms, regression gates, incident debugging, security controls, and operational ownership.
Additional Experience That May Be Helpful
  • Experience with HPC workflows, including Slurm, Kubernetes, multi-node GPU execution, containers, distributed launch, and shared clusters.
  • Experience automating enterprise workflows, developer platforms, IT operations, business systems, approvals, audit trails, or governed tool execution.
  • Experience with GPU profiling and performance analysis tools and trace-based workflows.
  • Experience with kernel authoring, tuning, or code generation using Triton, CUDA, HIP, MLIR, LLVM, or XLA-like flows.
  • Experience with LLM training, supervised fine-tuning, reinforcement learning, evaluation workflows, inference serving, model optimization, or agent evaluation.
Location

Finland or Sweden. Remote work arrangements may be considered for candidates located outside the Helsinki or Stockholm metropolitan areas.

 

#LI-MH3
#LI-HYBRID



Benefits offered are described:  AMD benefits at a glance.

 

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.

 

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD's “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.


 Apply on company website