Charger Logistics Logo

Charger Logistics

Site Reliability Engineer

Posted 2 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in India
Senior level
Remote
Hiring Remotely in India
Senior level
Maintain and improve a highly available logistics platform running .NET Core microservices on Kubernetes. Responsibilities include defining SLOs, building observability, managing Kubernetes and infrastructure as code, automating operations, supporting CI/CD, responding to incidents, tuning PostgreSQL, planning capacity, testing resilience, and reducing cloud costs. The role participates in on-call rotations and collaborates with development, DevOps, and QA teams.
The summary above was generated by AI

Charger Logistics Inc. is a leading asset-based transportation company with over 20 years of experience delivering innovative logistics solutions. We have evolved into a world-class transport provider and continue to expand across North America.

We're looking for a Site Reliability Engineer to keep our platform fast, stable and always available. Our systems run 24/7 to support trucking and logistics operations across North America. They're built on more than 700 .NET Core microservices running on Kubernetes with PostgreSQL. You'll work closely with our development, DevOps and QA teams in Canada to improve reliability, automate operations and respond to production issues.

Please note this is a Contract opportunity for potential to convert in a permanent role..

Responsibilities:

  • Keep our production platform reliable, scalable and highly available.
  • Define and track SLIs, SLOs and error budgets for critical services.
  • Build and improve monitoring, alerting, logging and tracing using Prometheus, Grafana, Loki, Jaeger and OpenTelemetry.
  • Manage and tune Kubernetes clusters and containerized workloads.
  • Automate repetitive operational work using Python, Go or Bash.
  • Build and maintain infrastructure as code with Terraform and Helm.
  • Support and improve CI/CD pipelines for safe, frequent deployments, including blue/green and canary rollouts.
  • Take part in an on-call rotation. Lead incident response, troubleshoot production issues and run blameless post-mortems.
  • Monitor and tune PostgreSQL performance, backups and high availability with the development team.
  • Plan capacity, test resilience and help reduce cloud costs.

Requirements
  • 6–8 years in SRE, DevOps or production infrastructure roles.
  • Strong hands-on experience with Kubernetes and Docker in production.
  • Experience with at least one major cloud platform (AWS, Azure or GCP).
  • Hands-on with monitoring and observability tools (Prometheus, Grafana, ELK/Loki, Jaeger or similar).
  • Infrastructure as code with Terraform, Helm or similar.
  • Scripting or coding skills in Python, Go or Bash.
  • Strong Linux and networking fundamentals (DNS, load balancing, TCP/IP, TLS).
  • Experience with CI/CD tools such as GitHub Actions, GitLab CI, Jenkins or Azure DevOps.
  • A solid understanding of microservices and distributed systems.
  • Experience handling production incidents and writing root-cause analyses.

Nice to have

  • Service mesh experience (Istio, Consul or Linkerd)
  • PostgreSQL administration and performance tuning
  • Familiarity with .NET Core applications in production
  • Messaging systems such as Kafka or RabbitMQ
  • Certifications such as CKA, CKAD, or AWS/Azure DevOps Engineer
  • Experience in logistics, transportation or another 24/7 operations environment is a plus

Benefits
  • Competitive Salary
  • Healthcare Benefit Package
  • Career Growth

Similar Jobs

9 Days Ago
Easy Apply
In-Office or Remote
Easy Apply
Senior level
Senior level
Cloud • Information Technology • Security • Software
Build and operate reliable, highly available cloud infrastructure across AWS and GCP. Develop observability, monitoring, automation, disaster recovery, Kubernetes, GitOps, and Terraform solutions. Participate in incident response and on-call rotations, define SLI/SLO practices, reduce operational toil through Python or Go, and improve resilience, cost visibility, and production maturity across microservices.
Top Skills: AlbAmazon EksArgo CdAWSAws Secrets ManagerClaude CodeCursorDatadogExternal Secrets OperatorGCPGithub ActionsGithub CopilotGitlab PipelinesGitopsGoHaproxyHashicorp VaultIamIstioKargoKubernetesNginxNlbPagerdutyPythonRoute 53TerraformVpc
2 Days Ago
Remote
Shri Bhrigukshetra, BLR, Uttar Pradesh, IND
Entry level
Entry level
Fintech • Analytics
Support and improve LSEG’s internal observability platform across production and non-production environments. Responsibilities include monitoring platform health, incident response, root cause analysis, service readiness, SLOs and SLIs, automation, GitOps workflows, documentation, and self-service enablement. The role uses metrics, logs, traces, dashboards, and telemetry standards to improve platform reliability, reduce manual work, and help engineering teams adopt consistent operational practices.
Top Skills: BigpandaCi/CdClickhouseCloud ComputingContainerizationCriblDatadogDistributed SystemsFlinkGitGitopsGrafanaInfrastructure As CodeLinuxNetworkingOpentelemetryRedis
9 Days Ago
Easy Apply
In-Office or Remote
Easy Apply
Senior level
Senior level
Cloud • Information Technology • Security • Software
Architects and operates highly available, multi-region cloud infrastructure, Kubernetes platforms, microservices, and authentication systems. Leads disaster recovery, failover automation, observability, SLOs, incident response, FinOps, infrastructure-as-code, and reliability improvements. Develops automation in Python or Go, manages GitOps deployments, mentors engineers, authors technical documentation, and participates in on-call rotations.
Top Skills: Amazon EksArgo CdAWSAws Secrets ManagerCert-ManagerClaude CodeCursorDatadogExternal Secrets OperatorGCPGithub CopilotGkeGoHaproxyIstioKargoKubernetesLinkerdNginxPagerdutyPythonTerraformVault

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account