Skip to content

DevOps Engineer - New Grad 2026

Build and maintain scalable infrastructure, CI/CD workflows, and cloud or datacenter automation for Cerebras’s AI software stack. The role requires a current university student or new graduate with software development experience and proficiency in Python, shell scripting, containers, Jenkins, and cloud platforms.

About the job

Responsibilities

  • Develop and maintain infrastructure required to build, test, operate, simulate, and evaluate the software stack.
  • Design efficient, scalable workflows for automating processes in the cloud and datacenter.
  • Collaborate with development and product management teams to monitor the quality and performance of software running on the Wafer Scale Engine.
  • Work across advanced hardware interfaces, low-level infrastructure, distributed systems, compilers, and machine learning frameworks.

Requirements

  • Enrolled in a university program pursuing a degree in Computer Science, Computer Engineering, or a related discipline.
  • Experience in software development environments.
  • Proficiency in Python, shell scripting, and Makefiles.
  • Strong end-to-end triage, debugging, and troubleshooting skills.
  • Experience with Jenkins and other CI/CD platforms.
  • Experience with Docker, Kubernetes, and container technology.
  • Experience building services on AWS or other cloud platforms at scale.

Nice-to-Haves

  • User interface experience.

Benefits

  • Opportunity to build a breakthrough AI platform beyond the constraints of GPUs.
  • Opportunities to publish and open-source AI research.
  • Work on a high-performance AI supercomputer.
  • Startup vitality with job stability.
  • A non-corporate work culture emphasizing individual beliefs, learning, growth, and support.

Skills

Python, Shell Scripting, Makefiles, Jenkins, CI/CD, Docker, Kubernetes, Containers, AWS, Cloud Platforms, Distributed Systems, Compilers, Machine Learning Frameworks

PointClickCare

PointClickCare

Mississauga, Canada

Junior Software Engineer
CA$70k+/yrHybridDevOps / SRE

Junior Site Reliability Engineer supporting production operations, observability, incident response, and automation for critical services. The role suits candidates with 0–2 years of experience, programming or scripting skills, and an interest in cloud infrastructure and distributed systems.

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Perplexity

Perplexity

San Francisco, CA
Member of Technical Staff
$220k+/yrRemote4+ YOEDevOps / SRE

Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.

Supabase

Supabase

Remote

Platform Engineer - Compute Capacity
No salary listedRemote5+ YOEDevOps / SRE

Platform engineer responsible for forecasting and automating compute capacity across regions, including reservations, fleet reconciliation, observability, and cost optimization. Requires 5+ years in infrastructure, SRE, platform, or capacity engineering plus production software and AWS EC2 experience.

Alpaca

Alpaca

Remote

Production Support Engineer
No salary listedRemote4+ YOEDevOps / SRE

Provides hands-on L2 technical escalation support for enterprise customers in the APAC region, troubleshooting distributed systems and APIs while leading root-cause analysis, support process improvements, and technical documentation. Requires 4+ years of support or escalation engineering experience.