Skip to content
ArmisArmis

AI Expert

Build and ship production-grade AI agents, assistants, and reusable skills for security-focused solutions. The role requires hands-on experience with agent frameworks, evaluation, Python, cloud integrations, and containerization, plus a bachelor's or master's degree or equivalent experience.

About the job

Responsibilities

  • Design, build, and bring to production AI models, agents, assistants, and reusable skills that solve real problems for Armis and its customers.
  • Turn ideas into working MVPs quickly, then partner with engineering to harden them into reliable, production-ready systems.
  • Build evaluations into every solution by defining quality criteria, measuring results, and maintaining rigorous standards.
  • Compose secure agentic workflows using secure harnesses, isolation, guardrails, and controls.
  • Select and integrate appropriate frameworks, tools, and models, and turn successful approaches into reusable patterns and best practices.
  • Participate in the Center of Excellence consulting team to support teams across the organization.
  • Document solutions, usage, operation, and evaluation practices.

Requirements

  • Experience building and shipping real AI solutions, including agents, assistants, copilots, or reusable skills.
  • Hands-on familiarity with modern agent frameworks and developer tooling, including leading provider SDKs and the SKILL.md skills standard.
  • Strong evaluation discipline, including building test sets, defining metrics, and measuring quality, safety, and reliability before and after deployment.
  • Proficiency in Python.
  • Experience integrating APIs, tools, and cloud services.
  • Experience with containerization.
  • Bachelor's or master's degree in Computer Science, Engineering, or a related field, or equivalent professional experience.

Preferred Qualifications

  • Experience in security or building AI for security use cases.

Compensation and Benefits

  • Comprehensive health benefits.
  • Discretionary time off.
  • Paid holidays, including monthly personal days.
  • Inclusive and diverse workplace.

Role Scope

  • This is an AI application development and systems-building role, not a machine-learning model-training or training-pipeline role.

Skills

Python, AI Agents, Agent Frameworks, Agent Sdks, Prompt Engineering, Retrieval, Evaluation Metrics, APIs, AWS, Azure, GCP, Docker, Kubernetes, Security Guardrails

Improbable

Improbable

Remote

AI Researcher
No salary listedRemoteAI Research

Conduct applied research on AI agents, designing experiments and evaluation systems to improve reliability, context retention, and multi-step task completion. The role requires strong AI/ML research, engineering, experimental design, and communication skills.

Wiz

Wiz

Tel Aviv, Israel

AI Security Researcher
No salary listedOn-site5+ YOEAI Research

Conducts deep technical research into cloud- and AI-native environments to identify novel risks and attack vectors, then translates findings into product capabilities with Product and Engineering teams. Requires at least five years of security research experience and strong scripting and telemetry-analysis skills.

AI Digest

AI Digest

Remote

Research Scientist - Member of Technical Staff
$150k+/yrRemoteAI Research

Conduct research on long-horizon, multi-agent AI behavior by designing agent environments, analyzing large-scale data, and running experiments. The role requires strong research judgment, rapid execution, independence, and familiarity with current AI developments.

AI Digest

AI Digest

Remote

Engineer - Member of Technical Staff
$150k+/yrRemoteAI Research

Build, optimize, and evaluate long-running and multi-agent AI systems, along with tools for monitoring and analyzing their real-world behavior. The role requires software engineering experience with coding agents, strong independence, and familiarity with current AI developments.

Vanta

Vanta

Remote

Senior Product Builder, Organizational Intelligence
$176k+/yrRemote5+ YOEAI Research

Build Vanta’s organizational intelligence layer by shipping prototypes, internal tools, and AI agent workflows that make cross-source data useful to EPD, GTM, and other teams. The role requires recent hands-on LLM product work, independent problem scoping, and strong judgment around AI quality, reliability, cost, and latency.