Teikametrics Logo

Teikametrics

Senior Site Reliability Engineer

Reposted 22 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in India
Mid level
Remote
Hiring Remotely in India
Mid level
The Senior Site Reliability Engineer will manage cloud infrastructure, develop internal DevOps tools, and enhance automation for application deployment and management.
The summary above was generated by AI
ABOUT THE COMPANY
Teikametrics is revolutionizing retail through our patented Artificial Retail Intelligence platform. Our proprietary orchestration layer functions as prompt intelligence, verticalizing AI for Amazon, Walmart, TikTok, and emerging marketplace use cases. For more information, visit www.teikametrics.com.
 
ABOUT THE ROLE
Teikametrics is looking for a Site Reliability Engineer at Bengaluru, India to help us build and maintain our cloud infrastructure for hosting Teikametrics applications and platforms, in addition to helping build internal DevOps tools and best practices required for efficient software development and deployment. 
 
This highly visible role will help us guide deployments using the latest technologies such as Docker, Kubernetes, and Terraform while having an enormous impact on our entire organization.
 
You will work within a DevOps model alongside product development teams, designing, deploying, and managing automation tools that increase predictability, improve efficiency, and reduce operational cost.
 
 
 
ABOUT THE TEAM
We currently host services and infrastructure across AWS in tandem with third party providers and must address the security and scaling challenges that come with this. Our daily role involves: 
  • Managing, and scaling web applications and data platforms

  • Building tools to implement devops and security best practices

  • Creating reusable and immutable infrastructure with Terraform

  • Continuously improve our infrastructure with monitoring, logging, and alerting

  • Developing authentication and gateway solutions for our infrastructure and applications

  • Investigating, identifying application issues and advising development teams on design, deployment and infrastructure choices.

  • Participating in on-call rotations, post-mortems and root cause analysis (RCA)

WHO YOU ARE

  • 5+ years of professional experience

  • Take full ownership of significant system components with responsibility for their reliability and performance

  • Manage Lifecycle of core project infrastructure, from design through to deployment, maintenance, and performance optimization 

  • Familiar with industry standards and devops best practices

  • Experience supporting the overall platform in on-call rotations.

  • Able to operate with minimal supervision

HOW YOU'LL SPEND YOUR TIME

  • Managing deployment infrastructure and automations including Github, CI/CD pipelines,and other deployment tooling

  • Experience managing workflows and pipelines using tools like CircleCI, Argo Workflows etc

  • Cloud computing providers such as AWS

  • Hands-on experience with Kubernetes (EKS, GKE) or similar container orchestration platforms 

  • Infrastructure as code tools such as Terraform

  • Experience coding with at least one language such as Bash, Python required

  • Hands-on experience with authentication and authorization technologies required 

  • Automation experience of cloud environments

  • Containerization technologies and tools such as Docker

  • Monitoring tools such as Datadog, Opensearch, Sentry

WHAT CAN HELP YOU STAND OUT

  • Good at using AI agents and writing project specific standard guidelines

  • Experience operating data pipelines with Databricks, Kafka

  • Experience with Java, Javascript

  • Experience with managing infrastructure costs and budgets

  • Experience with databases like AWS RDS/Postgres

WE'VE GOT YOU COVERED

  • Every Teikametrics employee is eligible for company equity
  • Remote Work Flexibility - Work from home or from our offices, with flexible remote options
  • Broadband reimbursement 
  • Group Medical Insurance – Coverage of INR 7,50,000 per annum for a family 
  • Crèche benefit

Press Reference about Teika
Teikametrics’ Marketplace Optimization Platform, Announces Artificial Retail Intelligence (ARI): An AI-Powered Tool Designed to Drive Cross-Marketplace Success
 
The job description is representative of typical duties and responsibilities for the position and is not all-inclusive. Other duties and responsibilities may be assigned in accordance with business needs. We are proud to be an equal opportunity employer. A background check will be conducted after a conditional offer of employment is extended. #LI-Remote
Beware of Recruitment Scams
Teikametrics will never ask you to communicate via WhatsApp, Telegram, or other messaging apps during our hiring process. We do not request payment, financial information, or purchases of equipment during recruitment. All legitimate job openings are posted on our official careers page at teikametrics.com/careers, and our recruiters will only contact you through LinkedIn or official company email addresses (@teikametrics.com).
 
If you're unsure about a communication you've received, please contact us directly at [email protected] to verify.

Similar Jobs

6 Days Ago
Remote or Hybrid
Senior level
Senior level
Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Lead reliability, scalability, observability, automation, infrastructure, disaster recovery, and security initiatives for Crunchyroll’s cloud-native data platforms. Establish SRE practices including SLIs, SLOs, error budgets, incident management, and postmortems. Operate Kubernetes and GCP environments, implement Infrastructure as Code, optimize capacity and performance, and drive vulnerability remediation, penetration-testing support, and cloud platform security.
Top Skills: Ci/CdDatadogGCPGoGrafanaIdentity And Access ManagementInfrastructure As CodeJavaKubernetesLinuxOpentelemetryOwasp Top 10PrometheusPythonShellTerraform
5 Days Ago
Remote
India
Senior level
Senior level
Information Technology • Productivity • Software • Manufacturing
Own reliability for major AWS production domains by defining SLOs, building observability and automation, managing capacity and self-healing, and leading complex incident response. Design Terraform modules and progressive delivery pipelines, operate ECS, EKS, Lambda, and PostgreSQL workloads, and reduce operational toil. Establish security and compliance controls, apply governed AI to operations, mentor SRE engineers, and standardize reliability practices across global teams.
Top Skills: AWSCi/CdCloudwatchDnsDockerEcs FargateEksGenerative AiGithub ActionsIamIso 27001KubernetesLambdaLlmsOpenobserveOpentelemetryPagerdutyPythonRds PostgresqlSoc 2TerraformVpc
20 Days Ago
In-Office or Remote
India
Senior level
Senior level
Cloud • Security • Software • Cybersecurity
Oversee, scale, and optimize high-density AI hardware infrastructure across regional data centers. Build Python automation and infrastructure-as-code tooling, integrate incident workflows, develop telemetry pipelines and monitoring dashboards, and improve reliability across private cloud, bare-metal, and virtualized environments. Lead on-call incident response, runbooks, post-mortems, service rollouts, vendor coordination, and field technician activities while driving uptime, performance, and operational readiness.
Top Skills: Ai-Based Anomaly DetectionApi IntegrationsBare-Metal InfrastructureBgpGrafanaInfrastructure As CodeIpv4Ipv6LlmsLokiOpentelemetryPagerdutyPrivate CloudPrometheusPythonSlackTelemetry PipelinesVirtualization

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account