Skip to content

Jobs

Anduril

Robot type
Drone · Defense
Location
Costa MesaCaliforniaUSA
Job type
Software
Posted
Aug 14, 2026
Salary
$166,000–$250,000 a year
Full-time

Senior Site Reliability Engineer - Undersea Dominance

Job description

Anduril Maritime delivers platforms, systems, and integrated effects in the maritime domain. Our autonomous vehicles (sub-surface and surface) are the cornerstone of these capabilities, and we continually strive to push the boundaries of the possible in terms of endurance, autonomy and mission capability. The Maritime team develops and maintains core products and payloads, and adapts and applies those products to serve a wide variety of defense, IC and commercial customers in US and international markets.

As a Senior Site Reliability Engineer on the Undersea Dominance team, you will build and operate the infrastructure that keeps our operational and production systems running at full speed.

Job responsibilities

  • Build and Manage CI/CD Pipelines: Develop and maintain CI/CD pipelines using tools like GitHub Actions and Jfrog Artifactory to ensure seamless integration and deployment of machine learning models and applications.
  • Infrastructure as Code (IaC): Utilize Terraform and Ansible to automate infrastructure provisioning and management on cloud platforms such as Azure, AWS, or Google Cloud Platform (GCP).
  • Containerization and Orchestration: Implement containerization solutions with Docker and manage container orchestration using Kubernetes to ensure reliable deployment and scaling of applications.
  • Monitoring and Logging: Establish comprehensive monitoring and logging solutions using tools like ELK Stack (Elasticsearch, Logstash, Kibana), Prometheus, and Grafana to ensure the smooth operation of deployment…
  • Collaborate with Cross-Functional Teams: Work closely with development, data science, and operations teams to foster collaboration and ensure the efficient and effective deployment of machine learning models.

Job requirements

  • Advanced proficiency in programming languages (Python for scripting and integration).
  • Experience with CI/CD tools like GitHub Actions, Jfrog Artifactory, Git, and CircleCI.
  • Proficiency with IaC tools (Terraform, Ansible).
  • Experience with cloud platforms (Azure, AWS, GCP).
  • Proficiency in containerization (Docker) and container orchestration (Kubernetes).
  • Knowledge of model registries and feature stores (e.g., MLflow, Kubeflow).
  • Experience with logging and monitoring tools ( Prometheus, Grafana).
  • Understanding of parallel computing frameworks (CUDA, OpenCL).

Similar jobs