Arrow Electronics, Inc. Logo

Arrow Electronics, Inc.

Senior Site Reliability Engineer (SRE)

Posted 22 Days Ago
Be an Early Applicant
In-Office or Remote
2 Locations
Senior level
In-Office or Remote
2 Locations
Senior level
Build, automate, and operate a global cloud platform: develop automation in Go/Python, manage large-scale EKS clusters (Karpenter), author Terraform and Helm IaC, lead incident response and post-mortems, define SLIs/SLOs, implement observability (Datadog/Prometheus/Grafana) and PagerDuty on-call, and develop secure self-service tools to meet SOC2. Night-shift role based in Ahmedabad, India.
The summary above was generated by AI
Position:Senior Site Reliability Engineer (SRE)

Job Description:

The Senior Site Reliability Engineer (SRE) will build, automate, and operate Hanwha Vision's global cloud platform. We treat operations as a software engineering problem. Your goal is to eliminate manual tasks and replace them with stable, automated systems.

 

 

Key Responsibilities

  • Automation: Write clean code (Go or Python) to automate cloud operations and deployment pipelines.
  • Kubernetes Engineering: Build and manage large-scale Amazon EKS clusters, including networking, security, and scaling (using Karpenter).
  • Infrastructure as Code: Write and maintain modular Terraform and Helm templates.
  • Incident Management: Lead troubleshooting for critical system outages. Write clear, blameless post-mortems to prevent issues from happening again.
  • Observability: Define SLIs/SLOs. Set up monitoring, logging, and on-call alerts using Datadog and PagerDuty.
  • Security & Compliance: Build secure, self-service tools (like automated JIT AWS access) to meet SOC2 requirements.

 

Required Technical Skills

  • 10+ years of professional experience in SRE, DevOps, or Systems/Infrastructure Engineering.
  • Coding: Strong programming skills in Go or Python.
  • Containers: Production experience running and scaling Kubernetes (EKS preferred).
  • IaC: Expert-level knowledge of Terraform.
  • Monitoring: Experience with Datadog, Prometheus, Grafana, or PagerDuty.
  • Database/Messaging (Preferred): Basic understanding of DynamoDB, Kakfa/MSK, or caching layers.
  • Languages: Strong written English proficiency is required to collaborate with global teams.

Certification : Preferred - AWS certified Solution Architect, AWS certified DevOps.
Remarks- ​Ready to work on Night shift only.

Location:IN-GJ-Ahmedabad, India-Ognaj (eInfochips)

Time Type:Full time

Job Category:Engineering Services

Similar Jobs

12 Days Ago
Remote
India
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Software • Cybersecurity
Design, build, operate, and scale cloud-native infrastructure and Kubernetes clusters across major clouds. Implement CI/CD, observability, IaC, SLIs/SLOs, automate with Go/Python, participate in 24x7 on-call, and contribute to open-source and technical knowledge sharing.
Top Skills: AiopsAmazon EksArgo CdAzure AksCloudnativepgGithub ActionsGitlab Ci/CdGoGoogle GkeGpuGrafanaIstioJenkinsKubernetesLokiMimirNode.jsOpenshiftOpentelemetryPrometheusPulumiPythonTerraformThanos
Yesterday
In-Office or Remote
India
Senior level
Senior level
Cloud • Security • Software • Cybersecurity
Lead reliability and performance investigations across Akamai's global edge and media/web delivery platforms. Troubleshoot distributed systems across application, network, platform, and OS layers; design observability (SLIs/SLOs, telemetry, dashboards, alerts); analyze performance and bottlenecks; develop automation, internal tools, and AI-assisted diagnostics; partner with engineering and operations on scalable fixes; and provide escalation and off-hours support during critical incidents.
Top Skills: Ai ModelsAlertsBashCachingDashboardsDnsGoHttp/HttpsLinuxProxiesPythonSlisSlosSQLTcp/IpTelemetryTlsUnix
14 Days Ago
In-Office or Remote
India
Senior level
Senior level
Cloud • Security • Software • Cybersecurity
Design, implement, and maintain reliable, scalable infrastructure for large distributed content delivery systems. Define and measure SLIs/SLOs, monitor availability and performance, troubleshoot incidents, and implement corrective actions. Develop automation to reduce manual work, participate in design reviews, and collaborate with product and engineering teams to improve system reliability and performance.
Top Skills: AdbmsBashCloud ComputingDatadogGrafanaJavaScriptOracle SqlPrometheusPythonUnix/Linux

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account