SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Salary: USD 101,000 - 161,000 / annual
Arista Networks is seeking a Senior Site Reliability Engineer to join the FedRAMP CloudVision-as-a-Service (CVaaS) SRE team. You will be responsible for operating and scaling Arista's global CloudVision service fleet, ensuring reliability, stability, and scalability of a Kubernetes-native enterprise network management and streaming telemetry SaaS platform.
In this role, you will drive, develop, and lead projects across multiple areas including data platform (NetDL) architecture and performance, capacity planning, autoscaling, disaster recovery, observability, CI/CD and change management, service network architecture, cost optimization, and infrastructure/cloud-first application security—all with a FedRAMP compliance perspective. You will join the on-call rotation split between West and East Coast timezones.
The CloudVision stack is built entirely on Kubernetes and runs on Google Cloud Platform (GCP) with Google Kubernetes Engine (GKE). The technical stack includes Golang, Python, Ansible/Pulumi, and Bash. You will develop, operate, and work with various database systems both on Kubernetes and leveraging managed database products. You will integrate with numerous open-source software projects that power the microservices stack and monitoring infrastructure.
Arista Networks is an industry leader in data-driven, client-to-cloud networking for large data center, campus, and routing environments. The company is known for innovation in cloud computing, artificial intelligence, and software-defined networking, and has received awards for Best Engineering Team, Best Company for Diversity, Compensation, and Work-Life Balance.
REQUIREMENTS:
- BS or MS in Computer Science or related technical field, or equivalent practical experience
- 5+ years of software engineering experience
- U.S. Citizenship required
- Experience managing FedRAMP SaaS environments or other highly regulated systems
- Proven experience developing or managing deployments of distributed database systems or large-scale SaaS applications
- Proficiency in Python, Go (Golang), or other programming languages
- Strong scripting skills in Bash or other scripting languages
- Familiarity with GCP and GKE preferred