Kaseya Logo

Kaseya

Sr. DevOps Engineer

Posted 9 Days Ago
Be an Early Applicant
In-Office
Pune, Maharashtra
Senior level
In-Office
Pune, Maharashtra
Senior level
Own the reliability, scalability, security, and performance of large-scale Linux infrastructure supporting backup and disaster recovery services. Manage Kubernetes, infrastructure as code, CI/CD pipelines, monitoring, observability, incident response, and production troubleshooting. Automate operational tasks with Shell, Python, or Go; implement security controls; and collaborate with engineering and SRE teams to improve deployment processes and platform reliability.
The summary above was generated by AI

About Kaseya

Kaseya is the leading provider of AI-powered IT management and cybersecurity software, serving Managed Service Providers (MSPs) and internal IT organizations worldwide. Our comprehensive platform helps organizations efficiently manage, secure, and automate their IT environments, driving operational efficiency and long-term business success.

Backed by Insight Partners, a leading global software investor, Kaseya has experienced sustained double-digit growth and continues to expand its global footprint. Today, Kaseya supports customers in more than 20 countries and manages over 15 million endpoints worldwide.

Founded in 2000, Kaseya was built by builders - and we're still building. We look for people who create rather than wait, who see a hard problem and lean in, and who treat challenges as raw material. At Kaseya, everyone plays a role in shaping the future of IT: whether you're in engineering, product, sales, marketing, customer support, or operations, your work helps protect, defend, and optimize IT environments across the globe.

We're building teams that grow, perform, and make an impact. If you're driven by the itch to make things better - a product, a process, a career - you'll fit right in.

At Kaseya, we don't just raise the bar. We build it.


Senior DevOps Engineer

Experience: 8–12 Years
Location: Pune (Hybrid/Onsite)

About the Role

We are looking for an experienced Senior DevOps Engineer with deep expertise in Linux Administration to join our Backup Platform Engineering team. In this role, you will own the reliability, scalability, and performance of large-scale Linux infrastructure that powers our next-generation backup and disaster recovery platform. You will work closely with software engineering, SRE, and platform teams to automate infrastructure, improve operational excellence, and ensure highly available production environments.


Key Responsibilities

  • Design, build, and manage highly available Linux-based production infrastructure supporting mission-critical backup services.
  • Administer and optimize large-scale Linux environments, including performance tuning, kernel configuration, storage, networking, and system troubleshooting.
  • Manage Kubernetes clusters, ensuring reliability, scalability, security, and efficient resource utilization.
  • Build and maintain Infrastructure as Code (Terraform/Pulumi) following reusable and modular design principles.
  • Design and enhance CI/CD pipelines using GitHub Actions, Jenkins, ArgoCD, or similar tools.
  • Develop automation using Shell scripting, Python, or Go to eliminate manual operational tasks.
  • Implement monitoring, logging, and observability using Prometheus, Grafana, Datadog, or similar platforms.
  • Drive incident response, root cause analysis, postmortems, and continuous operational improvements.
  • Collaborate with development teams to improve deployment processes, platform reliability, and production readiness.
  • Implement infrastructure security best practices including RBAC, secrets management, vulnerability scanning, and audit logging.
  • Troubleshoot complex infrastructure, networking, storage, and Linux operating system issues in production environments.

Required Skills

  • 8–12 years of experience in DevOps, Platform Engineering, Linux Administration, or Site Reliability Engineering (SRE).
  • Strong expertise in Linux system administration, including:
    • Performance tuning
    • Kernel parameters
    • Storage & filesystem management
    • Process management
    • System troubleshooting
    • Networking fundamentals
  • Hands-on experience managing Kubernetes clusters in production.
  • Strong knowledge of Infrastructure as Code using Terraform or Pulumi.
  • Experience designing and maintaining CI/CD pipelines using GitHub Actions, Jenkins, ArgoCD, or equivalent.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, ELK, etc.
  • Strong scripting skills in Shell, Python, or Go for infrastructure automation.
  • Experience managing production incidents, on-call support, alert tuning, and operational excellence.
  • Strong understanding of:
    • DNS
    • Load Balancers
    • Firewalls  
    • VPCs  
    • Hybrid networking
  • Experience implementing security best practices including Vault, RBAC, secrets management, and audit logging.

Preferred Skills

  • Experience with OpenStack (Nova, Swift, Neutron, Cinder).
  • Experience managing workloads across AWS, Azure, GCP, and private cloud environments.
  • Exposure to large-scale distributed infrastructure (5,000+ nodes).
  • Experience with storage platforms, backup infrastructure, and disaster recovery concepts (RPO/RTO).
  • Knowledge of cost optimization (FinOps), storage tiering, and infrastructure capacity planning.
  • Experience with Chaos Engineering and resilience testing.
  • Familiarity with bare-metal provisioning technologies such as PXE, MaaS, or Ironic.
  • Ability to read and troubleshoot Go-based services and contribute to automation tooling.

Additional information
Kaseya provides equal employment opportunity to all employees and applicants without regard to race, religion, age, ancestry, gender, sex, sexual orientation, national origin, citizenship status, physical or mental disability, veteran status, marital status, or any other characteristic protected by applicable law.

Similar Jobs

20 Days Ago
Hybrid
Senior level
Senior level
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Lead the design, operation, and standardization of global cloud infrastructure across AWS, Azure, and GCP. Own large-scale Kubernetes platforms, service mesh, GitOps, infrastructure as code, CI/CD, observability, security, compliance, and on-call excellence. Drive multi-account AWS architecture, mentor engineers, influence cross-functional technical direction, and develop AI-assisted and agentic operational automation using tools such as Amazon Bedrock.
Top Skills: Amazon BedrockAmazon EksArgocdAWSAws App MeshAws OrganizationsAzureAzure AksCi/CdFluxGCPGitopsGoGrafanaHelmIamIstioKargoKubernetesLinkerdLinuxMtlsOpensearchPrometheusPythonService Control PoliciesShell ScriptingTerraform
8 Days Ago
Remote or Hybrid
India
Senior level
Senior level
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Design, build, and deploy event-driven automations to eliminate operational runbook toil. Operate and support production databases and streaming platforms, participate in on-call rotation, implement least-privilege IAM and auditability, and lead automation roadmap across data infrastructure.
Top Skills: AlertmanagerApache AirflowAws CloudtrailAws CloudwatchAws EventbridgeAws IamAws LambdaClaudeCursorDynamoDBElasticsearchGithub ActionsGitopsJavaScriptJenkinsKafkaKubernetesMySQLPostgresPrometheusPythonRedisTerraformTypescript
4 Days Ago
In-Office
Senior level
Senior level
Software
Designs and maintains automation for provisioning, deployment, configuration, patching, testing, and troubleshooting of distributed Hyperscale environments. Develops Python, Shell, and Ansible frameworks; manages infrastructure across bare-metal, virtualized, containerized, cloud, and on-premises platforms; and supports VMware, Hyper-V, Kubernetes, OpenShift, and PXE deployments. The role also performs root-cause analysis, validates software across platforms, improves engineering productivity, and resolves complex product and customer issues.
Top Skills: AnsibleAWSAzureBoto3C++CobblerCorosyncDockerFabricGCPGitGithub CopilotHaproxyHyper-VKubernetesLinuxOpenshiftParamikoPxePythonPyyamlRacknS3ShellVMwareXML

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account