OT Systems Engineer
Designs and operates highly available OT infrastructure for AI supercomputer campuses, including power, cooling, facility controls, and industrial software platforms. Requires at least three years of OT or ICS administration experience, strong systems and networking skills, and onsite availability in the Memphis/Southaven area.
About the job
Responsibilities
- Design, deploy, and enhance OT environments supporting cooling plants, electrical distribution, liquid-cooling loops, on-site generation, battery energy storage systems, and data hall systems across current and future sites.
- Install, configure, maintain, and support BMS, EPMS, SCADA, PLC, and internally developed HMI platforms.
- Integrate IT technologies to simplify management, improve security, and create scalable, highly available industrial environments.
- Deploy and maintain development, test, and staging environments for controlled changes to live critical systems.
- Support cluster bring-up, capacity expansions, commissioning, and production training.
- Perform upgrades and maintenance between critical operations, including evenings and weekends as needed.
- Monitor services and respond to incidents affecting power, cooling, and environmental-control availability and performance.
- Use automation tools and contribute to infrastructure-as-code and DevOps initiatives for ICS/OT environments.
- Collaborate with controls, mechanical, and electrical engineers on simulation and emulation environments.
- Maintain standards, architectures, design drawings, operational procedures, and other technical documentation.
- Collaborate with IT, security, controls, facilities, construction, power teams, vendors, and integrators.
- Configure and maintain OT systems in accordance with industry and cybersecurity standards, including the Purdue Model and IEC 62443.
Requirements
- 3+ years of experience in OT systems engineering or industrial control systems administration.
- Hands-on experience with multiple industry-standard controls software platforms and tools.
- Significant experience designing, deploying, supporting, and troubleshooting OT environments in high-reliability settings.
- Willingness to participate in an after-hours on-call rotation and work extended hours or weekends as needed.
- Willingness to travel up to 20% between Memphis, Southaven, and other sites.
- Ability to lift 30 lbs, work at heights and in plant/data hall environments, and drive with a valid license.
- Ability to work onsite in the Memphis, Tennessee / Southaven, Mississippi area.
Preferred Skills and Experience
- Experience with real-time systems, industrial control networks, or OT environments in data centers, power generation, semiconductor, energy, or similar industries.
- Experience with BMS, EPMS, SCADA, and PLC systems supporting large cooling plants, medium-voltage distribution, and mission-critical facilities.
- Experience with hyperconverged, rugged industrial edge, and distributed compute architectures.
- Working knowledge of BACnet, Modbus, OPC UA, MQTT, Ethernet/IP, DNP3, controls networks, and OT cybersecurity.
- Proficiency with Bash, PowerShell, Python, Puppet, Terraform, Ansible, configuration management, provisioning, infrastructure as code, and DevOps tools.
- Familiarity with Active Directory, multi-platform authentication, and OT identity environments.
- Systems administration experience with Windows and Linux servers, databases, and storage/backup.
- Network administration experience, OSI model knowledge, and industrial network segmentation, VRFs, MDFs, and IDFs.
Skills
Bms, Epms, Scada, Plc, Python, PowerShell, Terraform, Ansible, Puppet, Active Directory, Windows Server, Linux, Bacnet, Modbus, Opc Ua
Similar jobs
DevOps / SRE jobsBuild and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.
Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.