Senior Software Engineer - Streaming AI
Builds and operates distributed backend infrastructure that enables reliable, scalable AI inference and agents on real-time streaming data. The role requires production experience with large-scale systems, cloud platforms, networking, and statically typed languages.
About the job
Responsibilities
- Design, develop, and operate large-scale, high-performance infrastructure powering Confluent Cloud.
- Build foundational software that improves reliability, scalability, and efficiency across cloud environments.
- Address distributed-systems challenges including consensus algorithms, failover strategies, and resource allocation.
- Collaborate across Confluent to optimize infrastructure for real-time data-streaming use cases.
- Troubleshoot and improve system reliability, observability, and performance across AWS, Azure, and Google Cloud.
Requirements
- 2–5 years of industry experience designing, building, and supporting backend systems in production.
- Strong fundamentals in distributed systems, cloud infrastructure, and networking.
- Experience building and operating large-scale, highly available systems.
- Understanding of cloud platforms and their services, including AWS, Azure, or Google Cloud.
- Proficiency in Java, Scala, C++, Go, or another statically typed language.
- Strong problem-solving skills and the ability to work in a fast-paced environment.
- BS, MS, or PhD in computer science or a related field, or equivalent work experience.
Nice-to-haves
- Exposure to model serving, LLM or agent infrastructure, or streaming data systems.
- Experience building and operating platforms that serve AI reliably at scale; ML research or model-training experience is not required.
Skills
Distributed Systems, Cloud Infrastructure, Networking, AWS, Azure, GCP, Java, Scala, C++, Go, Consensus Algorithms, Failover Strategies, Resource Allocation, Model Serving, Streaming Data
Similar jobs
Backend Engineering jobsBuild and operate backend systems for large-scale data ingestion, workflow processing, and AI inference. The role requires 8+ years of backend experience, distributed-systems expertise, and proficiency with Python, TypeScript, or Go in production AWS environments.
Build large-scale backend systems for Databricks’ data and AI platform, spanning distributed processing, cloud infrastructure, governance, and application runtimes. The role requires 5+ years of production experience in Java, Scala, or C++, plus expertise in distributed systems, SaaS architectures, cloud technologies, security, and SQL.
Build and scale backend services for the dbt metadata platform, including discovery, catalog, lineage, and run history. The role requires 5+ years of software engineering experience, distributed systems expertise, strong backend development skills, and hands-on cloud and containerization experience.
Senior Software Engineer responsible for evolving GitLab’s authorization model across its Ruby on Rails monolith and next-generation policy engine. The role requires production Rails experience, authorization expertise, security judgment, API knowledge, and strong written communication.
Senior backend engineer building authentication and identity services across GitLab’s Rails monolith and GATE. The role requires professional Go or Ruby experience, authentication and authorization expertise, and the ability to solve complex security and scalability problems in a remote environment.