DevOps Engineer - New Grad 2026
Build and maintain scalable infrastructure, CI/CD workflows, and cloud or datacenter automation for Cerebras’s AI software stack. The role requires a current university student or new graduate with software development experience and proficiency in Python, shell scripting, containers, Jenkins, and cloud platforms.
About the job
Responsibilities
- Develop and maintain infrastructure required to build, test, operate, simulate, and evaluate the software stack.
- Design efficient, scalable workflows for automating processes in the cloud and datacenter.
- Collaborate with development and product management teams to monitor the quality and performance of software running on the Wafer Scale Engine.
- Work across advanced hardware interfaces, low-level infrastructure, distributed systems, compilers, and machine learning frameworks.
Requirements
- Enrolled in a university program pursuing a degree in Computer Science, Computer Engineering, or a related discipline.
- Experience in software development environments.
- Proficiency in Python, shell scripting, and Makefiles.
- Strong end-to-end triage, debugging, and troubleshooting skills.
- Experience with Jenkins and other CI/CD platforms.
- Experience with Docker, Kubernetes, and container technology.
- Experience building services on AWS or other cloud platforms at scale.
Nice-to-Haves
- User interface experience.
Benefits
- Opportunity to build a breakthrough AI platform beyond the constraints of GPUs.
- Opportunities to publish and open-source AI research.
- Work on a high-performance AI supercomputer.
- Startup vitality with job stability.
- A non-corporate work culture emphasizing individual beliefs, learning, growth, and support.
Skills
Python, Shell Scripting, Makefiles, Jenkins, CI/CD, Docker, Kubernetes, Containers, AWS, Cloud Platforms, Distributed Systems, Compilers, Machine Learning Frameworks
Similar jobs
DevOps / SRE jobsJunior Site Reliability Engineer supporting production operations, observability, incident response, and automation for critical services. The role suits candidates with 0–2 years of experience, programming or scripting skills, and an interest in cloud infrastructure and distributed systems.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.
Platform engineer responsible for forecasting and automating compute capacity across regions, including reservations, fleet reconciliation, observability, and cost optimization. Requires 5+ years in infrastructure, SRE, platform, or capacity engineering plus production software and AWS EC2 experience.
Provides hands-on L2 technical escalation support for enterprise customers in the APAC region, troubleshooting distributed systems and APIs while leading root-cause analysis, support process improvements, and technical documentation. Requires 4+ years of support or escalation engineering experience.