name

AI Infrastructure & Platform Operations Engineer

Remote, United States remote Entry Salary not listed
remote Technology & IT Curated
Sign in to apply Free account — we bring you straight back to this role.

About the role

Role Overview
We are building a European AI Infrastructure & Platform Operations team responsible for operating large-scale AI infrastructure environments powered by NVIDIA GPUs, high-performance networking, Kubernetes, and next-generation platform technologies.
What You Will Do
Monitor, operate, and support production AI infrastructure platforms. Investigate and resolve infrastructure, networking, hardware, and platform-related incidents.
Why It Might Be a Fit
Gain exposure to NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments. Help define how next-generation AI infrastructure is operated and supported.
Requirements
3+ years of experience in infrastructure operations, platform operations, network operations, site reliability engineering, cloud operations, datacenter operations, or related technical roles
Strong Linux administration and troubleshooting skills
Good understanding of networking concepts and experience diagnosing infrastructure-related issues
Working knowledge of Kubernetes in production environments
Experience supporting production infrastructure and services
Strong analytical and problem-solving skills
Experience working within structured operational and incident management processes
Excellent communication and collaboration skills
Ability to work within a shift-based operational environment
Benefits
Work with some of the most advanced AI infrastructure environments in production today
Gain exposure to NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments
Help define how next-generation AI infrastructure is operated and supported
Be part of a team shaping the future of AI-powered operations through k0rdent AI
Join a growing organisation investing heavily in AI infrastructure and platform services
Originally posted on Himalayas

Interview prep

Walk in with sharper answers.

Use this as a quick practice sheet before you speak with the employer.

Role
Technology & IT Administration Operations Infrastructure Platform Engineer remote

Likely questions

  1. Tell us about work you have done that is close to the AI Infrastructure & Platform Operations Engineer role.
  2. How would you approach your first 30 days at name?
  3. Which of Administration, Operations and Infrastructure have you used recently, and what did it help you achieve?
  4. Describe a time you solved a problem without waiting to be told exactly what to do.
  5. How do you stay organised and communicate clearly when working remotely?

Prepare before the call

  • A recent example that proves your experience with Administration, Operations and Infrastructure.
  • One short story with a problem, your action, and the result.
  • Two examples that show the strengths listed on your CV.
  • A clear reason why this role and company interest you.
  • Your availability, preferred work style, and salary expectations.

Ask them

  • What would success look like in the first 90 days?
  • What are the main problems this hire should help solve?
  • How does the team give feedback and measure good work?
  • What does a normal working week look like for this role?
Practice line

I am interested in the AI Infrastructure & Platform Operations Engineer role because I can bring practical experience in Administration, Operations and Infrastructure, learn the team quickly, and contribute to the outcomes name needs from this hire.

Related jobs.

More roles from this company or category.