Site Reliability Operations Analyst - Commercial
Site Reliability Operations Analyst streamlines workflows, stabilizes projects, and acts as first responder for Palantir deployments to free engineers for technical work. Requires 3+ years project management, travel willingness, and strong judgment under pressure.
About the job
Core Responsibilities
- Work on many different types of problems and challenges, supporting deployments at large global customers or traveling for new pilot projects.
- Act as first responders when issues arise, making initial fixes and exhausting options before seeking help.
- Craft and implement processes to reduce friction and enable team members to focus on their strengths.
- Think creatively, work collaboratively, and do whatever it takes to get the job done.
What We Value
- Extraordinary judgment and composure in high-pressure situations.
- Creative project management with lightweight frameworks for rapid iteration and low-overhead customer updates.
- Proven track record of developing effective, collaborative customer relationships.
- Meticulous attention to detail, maintaining accurate records and tracking key project metrics.
- Enthusiasm for on-site customer work or supporting internal projects and senior leadership.
What We Require
- Ability to travel 25-75%, varies by location and team.
- 3+ years of project/program management experience, preferably in a fast-paced or dynamic environment.
Skills
Project Management, Process Optimization, Jira, Customer Relationship Management, Metrics Tracking, Incident Response, Travel Readiness
Similar jobs
DevOps / SRE jobsDesigns and operates foundational developer-infrastructure services for CI, builds, deployments, and testing. The role requires senior-level systems engineering, end-to-end service ownership, and cross-functional technical leadership.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.