Arista Networks Logo

Arista Networks

Senior Site Reliability Engineer - CloudVision

Posted 6 Days Ago
Be an Early Applicant
Hybrid
Dublin, IRL
Senior level
Hybrid
Dublin, IRL
Senior level
Build, deploy, and operate scalable, reliable, observable, and secure production systems. Develop infrastructure automation, monitoring, alerting, incident response, and staged deployment processes. Diagnose infrastructure issues, conduct postmortems, optimize platforms, manage maintenance windows, and collaborate with software engineering and external support teams. The role requires strong Linux administration, infrastructure-as-code, production operations, troubleshooting, and automation skills, with experience in cloud platforms, Kubernetes, Docker, monitoring, CI/CD, and distributed systems preferred.
The summary above was generated by AI
Company Description

Arista Networks is an industry leader in data-driven, client-to-cloud networking for large data center, campus and routing environments. Arista is a well-established and profitable company with over $8 billion in revenue. Arista’s award-winning platforms, ranging in Ethernet speeds up to 800G bits per second, redefine scalability, agility, and resilience.  Arista is a founding member of the Ultra Ethernet consortium. We have shipped over 20 million cloud networking ports worldwide with CloudVision and EOS, an advanced network operating system. Arista is committed to open standards, and its products are available worldwide directly and through partners.

At Arista, we value the diversity of thought and perspectives each employee brings. We believe fostering an inclusive environment where individuals from various backgrounds and experiences feel welcome is essential for driving creativity and innovation.

Our commitment to excellence has earned us several prestigious awards, such as the Great Place to Work Survey for Best Engineering Team and Best Company for Diversity, Compensation, and Work-Life Balance. At Arista, we take pride in our track record of success and strive to maintain the highest quality and performance standards in everything we do.

Job Description

Who You'll Work With

We are seeking an experienced and analytically-minded Site Reliability Engineer to join our organisation on a permanent, remote basis from Ireland. In this role, you will be instrumental in building, deploying, and operating critical production systems with a steadfast commitment to scalability, reliability, observability, and security. You will work collaboratively with cross-functional teams to ensure our infrastructure remains resilient, efficient, and future-ready. This is an excellent opportunity for a detail-oriented professional who thrives in a dynamic environment and is passionate about solving complex infrastructure challenges.

What You'll Do

  • Design, build, and deploy production systems with a focus on scalability, reliability, observability, and performance, ensuring systems meet stringent security standards
  • Develop and maintain comprehensive automation solutions to eliminate toil and streamline operational efficiency across production environments
  • Proactively monitor production systems, establish intelligent alerting strategies, and implement automated incident response mechanisms to minimise downtime
  • Create and maintain detailed incident response runbooks; conduct thorough postmortem analyses following incidents to identify root causes and prevent recurrence
  • Collaborate with software engineering teams to identify and resolve infrastructural bottlenecks, designing innovative solutions that enhance product deployment workflows
  • Manage and optimise monitoring infrastructure using industry-standard tools, ensuring comprehensive visibility across all systems
  • Plan, communicate, and execute maintenance windows on production systems with minimal disruption to service availability
  • Triage platform and infrastructural issues with decisiveness and analytical rigour; engage with third-party vendors and support teams as required
  • Deploy new systems and updates in a staged, risk-managed manner, ensuring safe and incremental rollouts
  • Survey and adopt best practices in infrastructure and platform management to maintain secure, scalable, and fault-tolerant systems
  • Study the design and implementation details of open-source systems to enhance troubleshooting capabilities and accelerate issue resolution
  • Work transparently with stakeholders to communicate system status, planned maintenance, and infrastructure improvements

#LI-SZ1

#automation #Ansible #Terraform #observability #Prometheus #Grafana #cloud platforms #AWS #GCP #Azure #container #orchestration #Kubernetes #Docker #CI/CD #Jenkins #GitLab

Qualifications

**Essential Requirements:**

  • Bachelor's degree in Computer Science, Engineering, or equivalent professional experience (5+ years in a related infrastructure or systems role)
  • Proficiency in one or more programming languages: Go, Python, or bash shell scripting, with the ability to implement medium-complexity automation workflows
  • Strong knowledge of Linux or UNIX from both administration and debugging perspectives
  • Hands-on experience operating software systems, infrastructure, and complex applications at scale in production environments
  • Demonstrated expertise in infrastructure-as-code principles and practices
  • Strong problem-solving and software troubleshooting skills with a methodical, analytical approach
  • Experience with server provisioning, particularly from storage and networking perspectives
  • Proven ability to work collaboratively within cross-functional teams and communicate technical concepts clearly
  • Experience with incident response, postmortem analysis, and continuous improvement methodologies

**Desirable Skills and Experience:**

  • Experience with container orchestration platforms, particularly Kubernetes
  • Hands-on experience with Docker and virtualisation technologies
  • Proficiency in managing monitoring stacks, including Prometheus and Grafana
  • Experience with CI/CD systems such as GitLab tools or Spinnaker
  • Knowledge of infrastructure-as-code frameworks, particularly Terraform
  • Experience managing databases such as PostgreSQL or equivalent relational database management systems
  • Experience with artifact repositories and Docker registries
  • Familiarity with cloud platforms (Google Cloud Platform, Amazon Web Services, or Microsoft Azure)
  • Understanding of distributed systems architecture and principles
  • Experience with performance tuning and system optimisation
  • Knowledge of security best practices in infrastructure and systems design
  • On-call support experience and comfort with incident response responsibilities

Additional Information

Arista stands out as an engineering-centric company. Our leadership, including founders and engineering managers, are all engineers who understand sound software engineering principles and the importance of doing things right.

We hire globally into our diverse team. At Arista, engineers have complete ownership of their projects. Our management structure is flat and streamlined, and software engineering is led by those who understand it best. We prioritize the development and utilization of test automation tools.

Our engineers have access to every part of the company, providing opportunities to work across various domains. Arista is headquartered in Santa Clara, California, with development offices in Australia, Canada, India, Ireland, and the US. We consider all our R&D centers equal in stature.

Join us to shape the future of networking and be part of a culture that values invention, quality, respect, and fun.

Arista Networks Dublin, Dublin, IRL Office

2 Georges Dock IFSC 1, Dublin, Ireland

Similar Jobs

An Hour Ago
Hybrid
Howth, Dublin, IRL
Senior level
Senior level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Senior Software Engineer responsible for building scalable backend systems in Java/Spring Boot, designing and implementing API and UI test automation, embedding quality/security into the SDLC, enhancing CI/CD, driving performance and reliability efforts, collaborating on architecture, and troubleshooting distributed systems.
Top Skills: BambooGitJavaJenkinsJSONJunitPlaywrightPostmanRestassuredRestful ServicesSeleniumSonarSpring BootTestngXML
An Hour Ago
Hybrid
Blackrock, Dublin, IRL
Senior level
Senior level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Lead end-to-end development of high-performance Java services for Mastercard’s distributed, in-memory fraud decisioning platform. Design scalable APIs and resilient distributed systems, troubleshoot multi-tier performance issues, coordinate releases, participate in on-call and incident management, improve automation and operational stability, and mentor engineers. Collaborate with product and business stakeholders on prioritization, demos, risks, and acceptance while driving engineering standards, secure development, testing, and continuous improvement.
Top Skills: AgileApache GeodeCi/CdCoherenceHazelcastHibernateJavaJbossJdk 17JenkinsJSONJunitLinuxMavenMessage Queuing SystemsPci-DssRedisRest ApisSafeScrumShell ScriptingSpringSpring BootSQLTomcat
An Hour Ago
Hybrid
Rathcoole, Dublin, IRL
Senior level
Senior level
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Own the full lifecycle of high-performance Java services and distributed caching systems supporting Mastercard’s real-time fraud decisioning platform. Responsibilities include architecture, development, testing, deployment, troubleshooting, release coordination, on-call support, incident management, performance optimization, stakeholder engagement, code reviews, mentoring, and automation. The role emphasizes scalable, secure, resilient systems, low latency, high availability, data consistency, and cross-team collaboration.
Top Skills: AgileApache GeodeCi/CdCoherenceHazelcastHibernateJavaJbossJdk 17JenkinsJSONJunitLinuxMavenMessage QueuesPci-DssRedisRestSafeScrumShell ScriptingSpringSpring BootSQLTomcat

What you need to know about the Dublin Tech Scene

From Bono and Oscar Wilde to today's tech leaders, Dublin has always attracted trailblazers, with more than 70,000 people working in the city's expanding digital sector. Continuing its legacy of drawing pioneers, the city is advancing rapidly. Ireland is now ranked as one of the top tech clusters in the region and the number one destination for digital companies, with the highest hiring intention of any region across all sectors.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account