Latest Data & Analytics jobs
Job results
Leads privacy and security engineering to secure Databricks' data platform, identifies infrastructure gaps, and drives strategy. Requires 9+ years in data security, 15+ years in distributed systems, and MS/PhD.
Leads security infrastructure engineering at Databricks, enhancing platform security, building scalable systems, and driving strategy. Requires 7+ years in data security, 10+ years in distributed systems, and MS/PhD.
Leads data security product strategy for Databricks Lakehouse, delivering controls for privacy, compliance, and access. Requires 7+ years PM experience and deep expertise in data security, privacy, encryption across multi-cloud environments.
Leads IAM and security engineering to secure Databricks' data platform, plugging infrastructure gaps and building scalable systems. Requires 9+ years in data security, 15+ years in distributed systems, and MS/PhD.
Engineering Manager leading IAM team to scale authentication, authorization, and identity services for Databricks' data and AI platform. Requires 5+ years leadership, deep IAM expertise, and experience with large-scale secure systems.
Engineering manager leading Spark Structured Streaming development, enhancing stream processing with advanced features and performance improvements. Requires 5+ years in big data/Spark, team leadership, and distributed systems expertise.
Leads technical direction and development of Unity Catalog's governance features for secure data and AI asset management at scale. Requires 15+ years in large-scale distributed systems, deep CS expertise, and strong leadership.
PhD intern researches and develops techniques to adapt LLMs and AI systems for enterprise domains, including method design, evaluation, and efficient post-training. Requires deep learning proficiency, PyTorch skills, and ongoing PhD studies.
Designs, implements, tests, and operates backend microservices for Databricks' large-scale data and AI platform using Scala/Java, Spark, Kafka, and cloud technologies. Requires 10+ years experience in distributed systems and SaaS platforms.
Designs and manages billing systems for Databricks products across major clouds, develops pricing primitives, enhances infrastructure reliability, and leads with AI for cost optimization. Requires BS in CS and proven distributed systems experience.
Staff Backend Software Engineer designs, builds, and operates scalable microservices for Databricks' data/AI platform using Scala/Java, Spark, Kafka, and cloud technologies. Requires 10+ years experience in large-scale distributed systems.
Build scalable backend infrastructure and products for Databricks platform, focusing on serverless, cloud core, partner ecosystems, notebooks, and application platforms. Requires 5+ years experience with Java/Scala/Golang/C++, distributed systems, cloud tech, and SaaS.
Build robust, scalable frontend experiences for Databricks' data and AI platform, focusing on user-centric workflows for data pipelines, visualization, and collaborative projects at massive scale. Requires 5+ years in JavaScript and modern frameworks.
Leads technical vision and execution for user activation features like onboarding and workspace experiences at Databricks. Requires 15+ years software engineering, 5+ years management, and full-stack expertise in scalable SaaS platforms.
Collaborates with customers and teams to architect big data solutions using Databricks platform for data engineering, science, and ML. Requires 5+ years experience, Python/SQL proficiency, cloud expertise, and distributed systems knowledge.
Develops distributed data systems like Apache Spark and Delta Lake at massive scale, ensuring high performance and reliability for exabyte-scale workloads. Requires 8+ years in Java/Scala/C++ and deep distributed systems expertise.
Designs and implements next-gen query engines and storage for Lakehouse architecture, focusing on query optimization, distributed execution, vectorized processing, and efficient storage. Requires 8+ years in database/distributed systems.
Fullstack engineer building scalable data platforms with user-friendly workflows for compute clusters, pipelines, and collaborative projects. Requires 5+ years in frontend (React) and backend (Node.js, Java, Python), plus cloud and distributed systems experience.
Designs, builds, and operates scalable backend microservices for Databricks' data and AI platform using Scala/Java, Spark, Kafka, and cloud technologies. Requires 10+ years experience in large-scale distributed systems.
Develop full-stack web applications for GenAI observability platform, focusing on intuitive UI/UX, scalable backend APIs, and performance optimization. Requires 5+ years in JavaScript frameworks and server-side technologies like Java/Python.
Leads technical vision and development of search quality systems, including ranking models, relevance evaluation, and hybrid retrieval for enterprise AI applications. Requires 10+ years experience in large-scale search and ML-driven relevance systems.
Designs, builds, and operates scalable microservices for Databricks' data and AI platform using Scala/Java, Spark, Kafka, and cloud technologies. Requires 10+ years experience in large-scale distributed systems.
Leads development of AI-powered systems for customer engagement using ML/LLMs to provide precise solutions and escalations. Requires 6+ years in large-scale distributed systems with strong product and ML expertise.
Pre-sales Solutions Architect for Healthcare/Life Sciences, partnering with Account Executives to drive technical strategy, ML/AI adoption, and data transformation for large enterprise accounts. Requires big data expertise, programming in Python/SQL/Scala, and bachelor's degree.
Leads compute fleet management across AWS, Azure, and GCP, optimizing billions of resources for peak performance, 99.99% availability, and 60%+ utilization. Requires deep distributed systems expertise and cross-team leadership for mission-critical infrastructure.
Designs and implements advanced query engines and storage systems for Lakehouse architecture, focusing on performance optimization across ETL, BI, and ML workloads. Requires 5+ years in database/distributed systems.
Designs and implements advanced query engines and storage systems for Lakehouse architecture, focusing on performance optimization across ETL, BI, and ML workloads. Requires 5+ years in database/distributed systems.
Build scalable backend infrastructure and products for Databricks' data and AI platform, focusing on resource management, distributed systems, and machine learning. Requires 5+ years experience with Java/Scala/C++, large-scale systems, and cloud technologies.
Build and extend scalable infrastructure for Databricks' data and AI platform, including multi-cloud systems and Kubernetes at massive scale. Requires 5+ years experience in Java/Scala/Go/C++/Python, distributed systems, and cloud technologies.
Designs and builds service-to-service communication systems, ingress control planes, and overload protection for Databricks' multi-cloud infrastructure. Requires 5+ years in distributed systems, proficiency in Java/Scala/Go/C++, and cloud/container experience.
Build AI/ML environment infrastructure enabling researchers to configure training and serving setups reliably. Requires 5+ years backend experience, strong Python/Scala/Java skills, and expertise in distributed systems and containerization.
Develop distributed data systems including Apache Spark and Delta Lake to handle big data workloads efficiently. Requires 5+ years in Java/Scala/C++ and expertise in distributed systems.
Senior engineer building distributed data systems like Apache Spark and Delta Lake to handle big data processing, ETL, and data science workloads. Requires 5+ years in Java/Scala/C++ and expertise in distributed systems.
Develops ML models for fraud/abuse detection and anomalous activity on Databricks platform. Analyzes security features, collaborates cross-functionally, and deploys production solutions. Requires 7+ years experience, MS in quantitative field, Python/SQL/Spark expertise.
Leads data science initiatives to inform business decisions, generate strategic insights for engineering priorities, and build production ML tooling. Requires 7+ years experience, strong Python/SQL/Spark skills, and MS/PhD in quantitative field.
Staff Data Scientist drives data-driven decisions through segmentation, recommendations, forecasting, and product analytics. Collaborates cross-functionally, mentors juniors, and requires 7+ years experience with Python/Scala, Spark, SQL, and MS/PhD.
Leads product management for Lakeflow in Data Engineering, owning vision, strategy, roadmap, and execution. Partners with teams to integrate across Databricks portfolio. Requires 5+ years PM experience and CS/engineering background.
Product Reliability Engineers own end-to-end service health for Palantir's critical platforms, combining on-call incident response with forward-looking work on observability, resilience, code improvements, and infrastructure migrations. Requires strong backend coding (Java), troubleshooting skills, and ownership in dynamic environments; US security clearance eligibility needed.
Frontend engineer building defense applications on Palantir's Foundry and AIP platforms. Owns user-facing features, works directly with customers, and leverages React, Maplibre, Three.js, and Redux to deliver mission-critical tools for the US military.
Build and operate data-heavy, customer-facing interfaces and workflows for the ClickPipes platform, working across frontend and backend systems. The role requires 5+ years of full-stack experience, deep React and TypeScript expertise, and production-quality testing and reliability practices.
Build and operate scalable, data-heavy user interfaces and customer-facing workflows for the ClickPipes platform, contributing across frontend and backend systems. The role requires 5+ years of full-stack experience, strong React and TypeScript expertise, and production-quality testing and operations practices.
Designs and implements cloud database features for Postgres platform, ensures operational excellence, performance, and observability for large-scale Postgres/Timescale instances on Kubernetes. Requires deep Postgres expertise, Golang, and experience managing stateful workloads at scale.
Senior backend engineer responsible for scaling ClickPipes’ real-time data onboarding platform and evolving its cloud infrastructure, deployment automation, and reliability. Requires 5+ years of backend experience plus strong Golang, Kubernetes, Terraform, and cloud-native architecture expertise.
Build and scale ClickPipes’ backend infrastructure and real-time data onboarding platform across cloud providers. The role requires 5+ years of backend experience, strong Go, Kubernetes, Terraform, distributed systems, and cloud-native architecture expertise.
负责在中国市场开发和维护大型企业客户,制定销售策略、拓展战略机会并推动收入增长。候选人需具备至少 5 年企业级软件销售经验,熟悉 AI、机器学习、大数据或基础设施软件,并能与高层客户进行价值沟通。
Leads complex engineering programs with cross-functional teams to deliver high-impact Snowflake platform features. Requires 5+ years experience in cloud technologies, strong technical judgment, and BS/MS/PhD in CS or equivalent.
The Solutions Engineer partners with sales teams and customers to demonstrate Snowflake’s data platform, translate business needs into use cases, and guide enterprise proofs of concept through implementation. The role requires strong presentation skills and experience with databases, data warehouses, ETL, analytics, cloud technologies, and SQL.
Optimizes the performance of ClickHouse’s core distributed database through query tuning, low-level systems work, regression testing, and production debugging. Requires professional C++ and Unix experience, strong database internals knowledge, and performance engineering expertise.
Develop and optimize the core ClickHouse database, focusing on distributed systems performance, query execution, testing, and production debugging. The role requires strong C++ and Linux experience, database internals knowledge, and collaboration with engineering and open-source communities.
Core software engineer responsible for optimizing ClickHouse database performance across query execution, distributed systems, caching, and low-level code. The role requires strong C++ and Unix experience, database internals knowledge, performance engineering expertise, and close collaboration with engineering and open-source communities.