Site Reliability Engineer (SRE/ DevOps) - Engineering Productivity

5 - 9 years

20.0 - 27.5 Lacs P.A.

Bengaluru

Posted:2 months ago| Platform: Naukri logo

Apply Now

Skills Required

UnixSANAutomationLinuxNetworkingMySQLShell scriptingDebuggingWindowsPython

Work Mode

Work from Office

Job Type

Full Time

Job Description

Who Youll Work With Arista Networks is looking for a skilled professional for our Engineering Productivity (EngProd) team to help maintain and support our rapidly expanding infrastructure and internal user base. The ideal candidate is someone who can wear many hats, is versatile and is enthusiastic about learning new technologies. As a part of the software engineering team, you will work with other team members to design, build and administer secure, scalable and fault-tolerant tools and infrastructure in a hybrid cloud environment. Working in the EngProd group, you will collaborate and work with other engineers to design, build, scale, and operate the systems used by Arista s product development teams. Thes systems are based on industry-standards, including Ansible, Artifactory, Gerrit, Jenkins, Kubernetes, Grafana, Spinnaker, MySQL, ElasticSearch, Google Cloud, Varnish, Perforce, Gerrit etc, 3rd party storage appliances, as well as internal systems developed from the ground-up to automate CI/CD, testing, analysis, and visualization. What Youll Do Build, deploy safely and incrementally and operate critical production systems with focus on scalability, reliability, observability, performance and security. Monitor, support and enhance developer experience across services. Build automation to remove toil and efficiently operate production systems. Proactively monitor, respond to, and enhance alerts and set up automated alert handling Create and maintain the incident response runbooks. Build and deploy new systems with scalability, reliability, and observability as primary requirements Triage platform/infrastructural issues and help Arista software engineers in their triages. Engage with 3rd party vendor support. Deploy new systems in a staged manner Write postmortem documents and build solutions to avoid incidents from repeating. Plan and communicate maintenance windows on production systems. Work with Arista s product development teams to identify infrastructural issues that are causing bottlenecks and limitations in their workflows. Design and implement solutions to resolve them. Survey and adopt best practices around infrastructure/platform to maintain secure, scalable and fault-tolerant systems. Implement solutions to scale the systems Implement fault-tolerance and performance to improve availability of the systems Study the design and sufficient implementation details of OSS systems for better triage and fix resolution. Essential to have all of the following skills At least BSc Computer Science or Engineering + 3 years experience, MS Computer Science or Engineering + 3 years experience, or equivalent work experience. Knowledge of one or more of

RecommendedJobs for You

Chennai, Pune, Delhi, Mumbai, Bengaluru, Hyderabad, Kolkata

Pune, Bengaluru, Mumbai (All Areas)

Chennai, Pune, Delhi, Mumbai, Bengaluru, Hyderabad, Kolkata

Bengaluru, Hyderabad, Mumbai (All Areas)

Hyderabad, Gurgaon, Mumbai (All Areas)