Bright Money Logo

Bright Money

SRE II

Posted 18 Days Ago
Be an Early Applicant
In-Office
Bangalore, Bengaluru Urban, Karnataka
Entry level
In-Office
Bangalore, Bengaluru Urban, Karnataka
Entry level
Own and scale AWS production infrastructure, improve CI/CD automation, lead incident response and RCA reporting, strengthen security and observability, manage cloud costs, and build cross-region disaster recovery and failover capabilities. The role requires operating Kubernetes clusters, Terraform infrastructure, monitoring and logging platforms, database and messaging systems, and cloud security controls.
The summary above was generated by AI

About Bright Money

Bright Money is a Consumer FinTech for middle-income consumers. It brings together SaaS products to help consumers build credit, manage their debt, and get access to suitable loans and credit cards. These are the major financial needs of middle-income consumers living paycheck-to-paycheck.
Bright is a profitable, India-built, U.S.-market Consumer FinTech, with revenue trajectory to $100M in the coming 12 months. The business model combines subscription revenue with financial service marketplace revenues for loans, credit cards, and credit-building tools. 

Bright has launched an AI money assistant and AI agentic UX flows for access to credit. It will be the first AI-driven Consumer FinTech for middle-income consumers.

The Bright product suite has potential for global markets beyond the US. The market size is 100M global middle-income consumers, with a revenue pool opportunity of $15B.

Bright is backed by 3 top global venture capital firms, and leading tech angel investors in the USA, Europe, and Asia.

Bright was founded in 2019 by a founding team from McKinsey’s Banking Practice (Petko Plachkov and Avi Patchava) and InMobi Data Scientists (Varun Modi and Avi Patchava). The core team is ex-InMobi Engineering and Data Sciences, recently joined by senior leaders from McKinsey. Bright is a data-driven, fast-paced execution culture that will build a Global Consumer FinTech, all functions operating from India.


What is the team & role about?


The Infrastructure/SRE team is responsible for building, managing, and scaling Bright Money’s cloud infrastructure to ensure our production systems are reliable, secure, scalable, and cost-efficient. The role involves owning AWS infrastructure, improving CI/CD and automation, strengthening observability and security, and ensuring the platform is prepared for incidents and disaster recovery.

You will work closely with engineering teams to operate production systems, automate infrastructure processes, improve system reliability, manage cloud costs, and build robust disaster recovery and failover capabilities.

What you will do?
  • Infrastructure Operations: Own and operate core production infrastructure on cloud platforms, ensuring high availability, observability, and scalability.

  • Incident Management: Lead incident response and author formal Root Cause Analysis (RCA) reports. Support the rollout of our new incident management platform and automated runbooks.

  • CI/CD & Automation: Design and optimize robust CI/CD pipelines. A major H2 goal is standardizing pipelines across all services following our Python upgrade and containerization tracks.

  • Security & Compliance: Implement infrastructure security compliance, including IAM roles, SCPs, and our upcoming Identity Platform (Teleport) rollout.

  • Observability: Maintain monitoring dashboards and alerting. Support the revamp of our VictoriaMetrics HA stack and ELK log optimization.

  • FinOps: Lead cloud cost analysis and resource tagging tracks to maintain efficient architecture.

  • Disaster Recovery: Lead the build-out of cross-region replicas and failover procedures to meet agreed RPO/RTO targets per service tier.

Ideal candidate profile looks like?Experience & Background:
  • Cloud Expertise: Strong experience with AWS (required) and Infrastructure-as-Code tools like Terraform.

  • Scripting: Proficiency in Python and Bash. Experience with the Django framework is a plus to support our internal tooling.

  • Observability Stack: Deep knowledge of Prometheus, Grafana, VictoriaMetrics, and ELK/OpenSearch.

  • Systems Knowledge: Strong analytical skills across database (RDS) and message queue (RabbitMQ) systems.

  • Security Mindset: Familiarity with SSO/IdP integration, secret management (Vault/AWS Secrets Manager), and vulnerability remediation.

  • Containerization: Hands-on experience with Kubernetes (EKS) and orchestrating multi-environment clusters. This is critical as we move toward full containerization of workloads in H2.

What Bright Offers?
  • A profitable sub-unicorn with U.S. market scale – liquidity potential at $1–3B valuations.

  • Real ownership: You run your own BU with an independent P&L.

  • Wealth creation: Equity at an inflection stage of scale.

  • A data-driven, high-cadence culture that combines consulting rigor with FinTech speed.

  • Location flexibility: 5 days from the Bangalore office in a week 

Similar Jobs

2 Hours Ago
Hybrid
Junior
Junior
Financial Services
Supports reliability and incident management for network services by troubleshooting infrastructure, building Python, Shell, and Ansible automation, reducing operational toil, and improving observability and SRE practices. The role works across routing, switching, firewalls, load balancers, proxies, SD-WAN, cloud infrastructure, and operating systems while applying authorized AI tools to incident analysis and reliability improvements.
Top Skills: AnsibleCisco AciCloud InfrastructureContinuous DeliveryContinuous IntegrationFirewallsLinuxLoad BalancersObservabilityProxiesPythonRouting And SwitchingSd-WanService Level ObjectivesShellSoftware-Defined NetworkingTelemetryWindows
2 Days Ago
In-Office
Mid level
Mid level
Artificial Intelligence • Machine Learning • Software • Analytics
Build, operate, and improve reliable production infrastructure across AWS and Kubernetes. Automate infrastructure with Terraform, develop CI/CD and GitOps workflows, and use Python, Go, or scripting to reduce operational toil. Monitor systems through metrics, logs, traces, and alerts; participate in on-call, incident response, root-cause analysis, and remediation. Partner with application, platform, and security teams to improve scalability, reliability, SLIs, SLOs, and operational efficiency.
Top Skills: Amazon CloudwatchAmazon EksArgocdAWSBashCi/CdDatadogGithub ActionsGitlab CiGitopsGoGrafanaJenkinsKubernetesLinuxPrometheusPythonTerraform
11 Days Ago
In-Office
Entry level
Entry level
Financial Services
Entry-level Site Reliability Engineer role supporting CME’s Globex trading platform. Responsibilities include learning observability, monitoring, alerting, automation, disaster recovery, resiliency testing, cloud migration, and production reliability practices. The engineer will write basic scripts, participate in blameless post-mortems, collaborate with product teams, and eventually join an on-call rotation under senior-engineer support. Strong foundational programming, problem-solving, communication, teamwork, and willingness to learn are required.
Top Skills: AWSAzureBashC++DockerGoogle Cloud Platform (Gcp)GrafanaHTTPJavaKubernetesLinuxPrometheusPythonSplunkTcp/Ip

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account