AI Fleet Management Solutions

Advanced Micro Devices, IncAustin, TX
6dHybrid

About The Position

At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. AMD is looking for an AI solutions systems Engineer who is passionate about complex and innovative AI solutions, AI infrastructure, fleet management and automation solutions. You will be a member of a core team of incredibly talented industry specialists and will work with the very latest hardware and software technology.

Requirements

  • Extensive experience building and operating distributed systems or large-scale infrastructure platforms
  • Strong proficiency in Python and JS based stacks (eg, Angular, React etc)
  • Experience designing and operating containerized systems (Docker, Kubernetes) and managing compute clusters (e.g., Slurm or similar schedulers)
  • Proven experience with public cloud platforms (AWS, Azure, or GCP) and/or hybrid or edge infrastructure
  • Strong knowledge of SQL databases
  • Deep understanding of Linux systems, administration, and scripting
  • General understanding in platform control, fleet management, IT and data center infrastructure
  • Familiar with AI dev tools such as Cursor or Claude
  • Demonstrated experience leading complex architectural initiatives across teams

Nice To Haves

  • Experience building AI-driven or agent-based systems in production environments
  • Contributions to open-source infrastructure projects
  • Experience with authentication and security
  • Experience with DevOps practices, CI/CD pipelines, and infrastructure-as-code tools (Terraform, Ansible, etc.)

Responsibilities

  • Work with AMD’s architecture specialists and hybrid cloud/edge developers to architect, develop and scale AI fleet management solutions spanning multiple global regions
  • Drive performance optimization and reliability based on industry best practices
  • Leverage AI to enhance developer and organizational productivity and solution capabilities
  • Influence roadmap and long-term architectural evolution of cutting-edge AI compute solutions

Benefits

  • AMD benefits at a glance.

Stand Out From the Crowd

Upload your resume and get instant feedback on how well it matches this job.

Upload and Match Resume

What This Job Offers

Job Type

Full-time

Career Level

Mid Level

Number of Employees

5,001-10,000 employees

© 2024 Teal Labs, Inc
Privacy PolicyTerms of Service