Anduril Industries

Senior Site Reliability Engineer

$191K–$287KFull-time · Costa Mesa, CA +1 more
✓ Verified live on the employer's own system · added 381 days ago
Save search
Senior

Skills & tools

DevopsCloud PlatformsPythonOperationsRoot Cause AnalysisRustC Plus PlusSecurity Clearance
Apply on company site ↗ See your fit → free

Full job description

For this unique opportunity, you'll be supporting both the Counter Intrusion team and the Air Defense team.

The Counter Intrusion team creates robotic systems that provide force protection capabilities by monitoring the perimeter of secure areas for approaching people, vehicles, and vessels. Leveraging advanced sensor fusion and autonomy, our products seamlessly render activity in the environment to Lattice's common operating picture.

The Air Defense team fields Anduril's most complex capabilities, deploying unique combinations of hardware and software tailored to different mission requirements and customer expectations. Our systems integration engineers internalize the nuances of each deployment, ensuring the technical success of the end-to-end solutions we ship.

We are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine or Costa Mesa. SREs work with external stakeholders to determine the technical direction of cloud deployments and deliver with speed through analysis, design and code. They are comfortable leading large, focused projects.

They lead in the development of Kubernetes cloud infrastructure, DevOps, CI/CD and improving the developer experience. You will be managing cloud deployments in AWS, Azure and on premise. This role emphasizes the continuous innovation in improving our cloud computing environments.

- Architect, deploy and maintain infrastructure with cloud providers and Kubernetes (EKS)

- Collaborate with multi-disciplined teams to define and execute on internal and external deployments

- Promote SRE best practices in system resilience, performance monitoring and high availability

- Design, develop, and deliver solutions using infrastructure as code with tools like Terraform and Python

- Develop and maintain CI/CD pipelines for automated deployment

- Build strong relationships with internal and external customers to identify technical solutions to their problems

- Improve Anduril's operational capabilities by improving our core product offering through root cause analysis and creating tooling capable of managing large scale deployments

- Lead the organization in building scalable, sustainable mechanisms to continue delivering to customers at the pace the business is scaling

- Technical expertise and demonstrated performance in one or more of the following areas: networking, cloud technologies, application development and/or cybersecurity

- Deep knowledge of the Kubernetes ecosystem (Docker, Helm, ArgoCD, Terraform)

- Experience in software languages such as Go, Python, Rust, or C++

- Experience performing data-driven root cause analysis on complex systems

- Demonstrated ability to train peers or customers on the operation of a product

- Eligible to obtain and maintain an active U.S. Secret security clearance

- Experience with managing Kubernetes clusters of hundreds of nodes

- Knowledge of performance improvement techniques, metrics and alerting

- Experience with KubeVirt, qemu, virtualization and hypervisor technologies

More jobs at Anduril Industries

Similar jobs near Costa Mesa, CA +1 more

Tell me when more Senior Site Reliability Engineer, Fleet Management jobs post near Costa Mesa, CA +1 more We re-check every listing against the employer’s own board — no résumé needed.

Search Senior Site Reliability Engineer jobs near Costa Mesa, CA +1 more → Browse all live jobs

This posting was published by Anduril Industries on their own careers system and is shown here with a direct link to apply there. Employers: for corrections or removal, contact jobs@veritahire.com.