yoinka

Senior Site Reliability Engineer (SRE)

Juniper Networks

Bengaluru Karntaka IndiaSeniorH-1B sponsor company
Sign in to applyVerified 1h ago
Location
Bengaluru Karntaka India
Work model
On-Site
Level
Senior
H-1B history
140 approvals (FY2023)
Posted
27d ago

Skills

AWSAnsibleCI/CDCassandraCloudFormationDockerElasticsearchFlinkGCPGitGitHub ActionsGoJenkinsKafkaKubernetesLinuxPostgreSQLPrometheusPythonRedisRustSparkTerraform

About this role

Senior Site Reliability Engineer (SRE) This role has been designed as ‘’Onsite’ with an expectation that you will primarily work from an HPE office.

Who We Are

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today’s complex world. Our culture thrives on finding new and better ways to accelerate what’s next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE.

Job Description

What You Will Do: Ensure high availability, reliability, and performance of large-scale cloud infrastructure across AWS and GCP environments. Operate and support infrastructure components and distributed data platforms such as Kubernetes, Kafka, Flink, Storm, and Spark . Manage and maintain databases including Cassandra, Elasticsearch, Redis, Postgres, and ArangoDB . Monitor systems, troubleshoot issues, and resolve production incidents across microservices and distributed systems. Collaborate closely with software engineering teams to debug and resolve complex production problems. Participate in 24x7 on-call rotation supporting multi-cloud production environments. Monitor system metrics, application performance, and infrastructure health using observability tools. Own the incident management lifecycle , including detection, mitigation, Root Cause Analysis (RCA), and post-incident reviews. Develop and maintain runbooks, automation, and operational processes to improve reliability and efficiency. Perform capacity planning using system usage and performance data. Drive SRE best practices, operational standards, and continuous improvement initiatives .  What You Need to Bring: Bachelor’s or Master’s degree in Computer Science, Information Systems, or a related field. 6–10+ years of experience in DevOps, Site Reliability Engineering, or cloud infrastructure roles. Strong hands-on experience with cloud platforms (AWS or GCP) including services like EC2/GCE, IAM, and object storage (S3/GCS). Experience with containerization and orchestration technologies , especially Docker and Kubernetes . Experience building and managing CI/CD pipelines using tools such as Jenkins, GitHub Actions, or GitLab . Experience with monitoring and observability tools such as Prometheus, CloudWatch, or Stackdriver . Strong understanding of Linux systems administration and configuration management tools like Ansible . Experience managing distributed systems and streaming platforms such as Kafka, Cassandra, Elasticsearch, Spark, Flink, or Storm . Strong automation and scripting skills using Python, Go, Rust, or Shell scripting . Experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation . Excellent analytical, troubleshooting, and problem-solving skills . Strong communication and collaboration skills with the ability to work with cross-functional teams. What We Can Offer You: Health & Wellbeing We strive to provide our team members and their loved ones with a comprehensive suite of benefits that supports their physical, financial and emotional wellbeing. Personal & Professional Development We also invest in your career because the better you are, the better we all are. We have specific programs catered to helping you reach any career goals you have — whether you want to become a knowledge expert in your field or apply your skills to another division. Unconditional Inclusion We are unconditionally inclusive in the way we work and celebrate individual uniqueness. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our

Senior Site Reliability Engineer (SRE) at Juniper Networks, Bengaluru Karntaka India | Yoinka