
200 open roles · Software Engineering
Site Reliability Engineer - SRE (Platform Software Team)
- Added to ZestAmigo
- Last seen on employer site
Where you can work
Hybrid
Bengaluru, Karnataka, India
View location wording from the posting
Bengaluru, KA, India
Role overview
Extracted from the posting. Read the employer’s full description below.
What they are looking for
- At least Bachelors in Computer Science or Engineering + 5 years’ experience, MS Computer Science or Engineering + 5 years’ experience, or equivalent work experience.
- Knowledge of one or more of Go, Python, bash shell scripting to be able to implement medium complexity automation workflows.
- Knowledge of Linux (or UNIX) from administration and debugging perspective
- Hand-on experience in deploying and managing Kubernetes on bare-metal infrastructure
- Experience in server provisioning with Ansible and Ansible Tower/AWX
Show all 15 itemsShow fewer items
- Experience with docker and virtualization technologies
- Experience with air-gapped systems
- Strong problem solving and software troubleshooting skills
- Experience with infrastructure-as-code.
- Desirable to have one/more of the following skills
- Experience managing databases - eg: MySQL or equivalent RDBMS etc
- Experience managing monitoring stack - Prometheus, Grafana etc
- Experience managing Artifactory, docker registry etc
- Experience managing CI/CD systems like GitLab CI/CD, Jenkins, Zuul etc
- Experience with infrastructure-as-code frameworks like Terraform
Employer description
Company Description
Arista Networks is an industry leader in data-driven, client-to-cloud networking for large data center, campus and routing environments. Arista is a well-established and profitable company with over $8 billion in revenue. Arista’s award-winning platforms, ranging in Ethernet speeds up to 800G bits per second, redefine scalability, agility, and resilience. Arista is a founding member of the Ultra Ethernet consortium. We have shipped over 20 million cloud networking ports worldwide with CloudVision and EOS, an advanced network operating system. Arista is committed to open standards, and its products are available worldwide directly and through partners.
At Arista, we value the diversity of thought and perspectives each employee brings. We believe fostering an inclusive environment where individuals from various backgrounds and experiences feel welcome is essential for driving creativity and innovation.
Our commitment to excellence has earned us several prestigious awards, such as the Great Place to Work Survey for Best Engineering Team and Best Company for Diversity, Compensation, and Work-Life Balance. At Arista, we take pride in our track record of success and strive to maintain the highest quality and performance standards in everything we do.
Job Description
Who You'll Work With
As a Site Reliability Engineer within the Platform Software Team, you will focus on developing and sustaining the fundamental infrastructure that supports our Manufacturing and Platform Diagnostics teams. This critical environment facilitates the validation of high-speed digital designs and ensures optimized manufacturing yields for Arista Network products, which serve the largest data centers across the computer networking sector
In this role, you will collaborate closely with a core team of Infrastructure and Tools Engineers. Serving as a bridge across multiple Arista engineering groups, you will develop systems and tools utilized within our HW Labs and Manufacturing sites to ensure that lab and manufacturing operations run smoothly.
What You’ll Do:
As a SRE, you’ll be responsible for the software infrastructure powering our HW Labs and Manufacturing Sites. This includes:
- Build, deploy safely and incrementally and operate critical production systems with focus on scalability, reliability, observability, performance and security.
- Monitor, support and enhance product deployment experience across services.
- Build automation to remove toil and efficiently operate production systems.
- Proactively monitor, respond to, and enhance alerts and set up automated alert handling
- Create and maintain the incident response runbooks.
- Build and deploy new systems with scalability, reliability, and observability as primary requirements
- Triage platform/infrastructural issues and help Arista software engineers in their triages. Engage with 3rd party vendor support.
- Deploy new systems in a staged manner
- Write postmortem documents and build solutions to avoid incidents from repeating.
- Plan and communicate maintenance windows on production systems.
- Work with Arista’s product development teams to identify infrastructural issues that are causing bottlenecks and limitations in their workflows. Design and implement solutions to resolve them.
- Survey and adopt best practices around infrastructure/platform to maintain secure, scalable and fault-tolerant systems.
- Implement solutions to scale the systems
- Implement fault-tolerance and performance to improve availability of the systems
- Study the design and sufficient implementation details of OSS systems for better triage and fix resolution
Qualifications
- At least Bachelors in Computer Science or Engineering + 5 years’ experience, MS Computer Science or Engineering + 5 years’ experience, or equivalent work experience.
- Knowledge of one or more of Go, Python, bash shell scripting to be able to implement medium complexity automation workflows.
- Knowledge of Linux (or UNIX) from administration and debugging perspective
- Hand-on experience in deploying and managing Kubernetes on bare-metal infrastructure
- Experience in server provisioning with Ansible and Ansible Tower/AWX
- Experience with docker and virtualization technologies
- Experience with air-gapped systems
- Strong problem solving and software troubleshooting skills
- Experience with infrastructure-as-code.
- Desirable to have one/more of the following skills
- Experience managing databases - eg: MySQL or equivalent RDBMS etc
- Experience managing monitoring stack - Prometheus, Grafana etc
- Experience managing Artifactory, docker registry etc
- Experience managing CI/CD systems like GitLab CI/CD, Jenkins, Zuul etc
- Experience with infrastructure-as-code frameworks like Terraform
Additional Information
Arista stands out as an engineering-centric company. Our leadership, including founders and engineering managers, are all engineers who understand sound software engineering principles and the importance of doing things right.
We hire globally into our diverse team. At Arista, engineers have complete ownership of their projects. Our management structure is flat and streamlined, and software engineering is led by those who understand it best. We prioritize the development and utilization of test automation tools.
Our engineers have access to every part of the company, providing opportunities to work across various domains. Arista is headquartered in Santa Clara, California, with development offices in Australia, Canada, India, Ireland, and the US. We consider all our R&D centers equal in stature.
Join us to shape the future of networking and be part of a culture that values invention, quality, respect, and fun.
Track this application
Keep your own notes. Only you can mark an application as sent.
Report a problem with this listing
Sign in to report this listing.
More at Arista Networks
Unspecified
Santa Clara, California, United States
Full-time
$175k to $200k USD
Posted 5 months ago
Source: SmartRecruiters
Unspecified
Mexico City, Mexico City, Mexico
Full-time
Salary unavailable
Posted 5 months ago
Source: SmartRecruiters
Unspecified
Santa Clara, California, United States
Full-time
$123k to $191k USD
Posted 5 months ago
Source: SmartRecruiters
Unspecified
Austin, Texas, United States
Full-time
Salary unavailable
Posted 5 months ago
Source: SmartRecruiters
Similar roles elsewhere
Hybrid
San Mateo, California, United States
Full-time
$170k to $200k USD / year
Posted 11 hours ago
Source: Ashby
Remote
Remote eligibility: United States
Full-time · Senior
$110k to $130k USD / year
Posted 1 month ago
Source: Workday
Remote
Remote eligibility: Spain
Full-time · Senior
Salary unavailable
Posted yesterday
Source: SmartRecruiters
Remote / Unspecified
Remote eligibility: Eligibility not captured
United States +1 more
Full-time
$100k to $125k USD / year
Posted yesterday
Source: Greenhouse