Moxie (joinmoxie.com) Logo

Moxie (joinmoxie.com)

Staff Platform Engineer (IC-4)

Reposted 7 Days Ago
In-Office or Remote
Hiring Remotely in India
Senior level
In-Office or Remote
Hiring Remotely in India
Senior level
Ownership of platform reliability and incident response across cloud infrastructure. Improve observability, CI/CD pipelines, deployments, developer experience, and operational tooling. Act as primary on-call platform expert, collaborate with product teams, and run postmortems to raise platform reliability.
The summary above was generated by AI

At Moxie, we empower ambitious aesthetic entrepreneurs to build profitable, independent practices—without burnout, overwhelm, or guesswork. In just a few years, we've grown from an idea to a global, remote-first team now supporting 700+ practices nationwide.

Our purpose is simple: to unlock sustainable success for aesthetic entrepreneurs, at every stage of their journey.

Staff Platform Engineer

Remote, Full-time
Location:
LATIN AMERICA Colombia, Brazil, Dominican Republic, or Chile - Fully Remote (Work from Home)
Working hours: Core overlap with 9 AM – 5 PM EST (flexible schedules between 7 AM – 8 PM EST)

About Moxie

Moxie empowers aesthetic industry professionals to become successful entrepreneurs. We provide a sophisticated SaaS platform that simplifies the operational complexities of running MedSpas, enabling nurses and medical professionals to launch, operate, and grow their businesses across the country.

Hundreds of customers rely on Moxie Suite to run their MedSpas end-to-end: scheduling, medical purchasing, payments and invoicing, bookkeeping, analytics, and more. The platform operates at real-world scale and reliability requirements, integrating with systems such as AWS, Vercel, Cloudflare, Stripe, Twilio, Datadog, and others.

We are a fast-growing company focused on building reliable infrastructure, strong developer experience, and operational excellence as we scale.

 
 
The Role

We are looking for a Staff Platform Engineer to help own and evolve the systems that enable Moxie engineers to ship safely, quickly, and reliably.

This role sits at the intersection of DevOps and SRE. Your most important responsibility will be incident handling — you'll be the primary owner of detecting, responding to, and resolving production issues across our infrastructure. During US business hours, you'll typically be the sole platform/infra expert on call, though developers will be available to support you as needed. Given this, the role can be demanding at times and requires availability beyond standard hours when incidents arise.

Beyond incident response, you'll work on cloud infrastructure, CI/CD pipelines, deployment workflows, local development environments, observability, and operational tooling. You'll partner closely with product engineering teams, but your primary responsibility is the health, reliability, and usability of the platform itself.

This is a hands-on individual contributor role with meaningful ownership, but no people management responsibilities.

Key ResponsibilitiesObservability & Reliability
  • Participate in incident response as needed

  • Own and improve monitoring, logging, and alerting using Datadog, AWS, Vercel and related tools

  • Ensure systems are observable and failure modes are well understood

  • Help teams learn from incidents through postmortems and follow-ups

  • Balance reliability with delivery speed through pragmatic SRE practices

Platform & Infrastructure
  • Own and operate core platform systems across AWS, GCP, Vercel, Github, and Cloudflare

  • Improve reliability, scalability, and security of production and non-production environments

  • Maintain and evolve infrastructure supporting multiple services and teams

CI/CD & Deployments
  • Own and improve CI/CD pipelines (GitHub Actions), focusing on speed, reliability, and clarity

  • Improve deployment workflows, rollbacks, and environment consistency

  • Reduce deployment-related risk and manual intervention

  • Partner with engineers and our QA team to improve release confidence and velocity

Developer Experience & Local Development
  • Improve local development environments and onboarding experience for engineers

  • Reduce friction in common workflows (setup, testing, debugging)

  • Maintain tooling and documentation that helps engineers move faster with confidence

Cross-Team Collaboration
  • Work closely with the product engineering team to understand platform pain points and improve local development experience.

  • Provide guidance and support on infrastructure, deployments, and operational best practices

  • Contribute to platform standards and shared tooling through collaboration, not mandates

QualificationsRequired
  • 5+ years of experience in platform, DevOps, or SRE-focused roles

  • Strong experience operating production systems on AWS (certifications strongly preferred), Vercel, and/or GCP

  • Experience building and maintaining CI/CD pipelines (GitHub Actions or similar)

  • Strong understanding of cloud networking, security fundamentals, and IAM

  • Experience with observability tooling (Datadog preferred)
    Ability to troubleshoot production issues calmly and systematically

  • Excellent written and verbal communication skills in English (C1 or higher).

Nice to Have
  • Experience with Cloudflare (DNS, WAF, edge configuration)

  • Experience with Agentic and LLM based tooling and automations (Cursor, Codex, Claude Code, etc)

  • Experience improving local development tooling (ie Docker, Husky, bash)

  • Familiarity with infrastructure-as-code (Terraform or similar)

  • Experience supporting regulated or compliance-sensitive environment

  • Experience working with PII, PHI, and sensitive data systems in general.

Our Stack
  • Cloud & Infrastructure: AWS ECS/Fargate, GCP, and Vercel

  • Edge & Security: Cloudflare

  • CI/CD: GitHub Actions

  • Observability: Datadog

  • Database: AWS RDS (postgres)

  • Version Control: Git, GitHub

  • AI / LLM Tooling: Claude Code, Gemini, Cursor, CodeRabbit, Glean, Codex

  • Backend Services: Python, Django (operational ownership, not feature dev)

At Moxie, we believe in creating a workplace where everyone feels valued, trusted, and included. Our team lives by our values: act as owners, give more than we take, move with speed and care, and simplify and learn every day.

We welcome people of all backgrounds, experiences, and perspectives to apply. If you require any accommodations to fully participate in the interview process, please let us know, we’re happy to assist.

Similar Jobs

3 Hours Ago
Easy Apply
Remote or Hybrid
Easy Apply
Senior level
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Design and build Samsara’s cloud governance platform, including IAM guardrails, security controls, compliance automation, and software supply-chain security tooling across AWS and GCP. Develop reliable distributed systems, secure CI/CD patterns, and SAST/DAST capabilities. Contribute to technical roadmaps, collaborate across infrastructure, security, compliance, and engineering teams, and communicate design tradeoffs while mentoring teammates through reviews and pairing.
Top Skills: AWSCi/CdCloud InfrastructureDastDistributed SystemsGCPGoIamPythonSastTerraform
4 Hours Ago
Easy Apply
Remote
Easy Apply
Senior level
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Lead frontend engineering for Affirm’s consumer web experiences using React and TypeScript. Own quarterly team goals, deliver major product features, modernize web architecture and design systems, and develop tools such as codemods and AI agent skills. Collaborate with product, design, analytics, and engineering teams; establish quality standards; monitor availability and metrics; support on-call operations; resolve technical issues; and mentor engineers.
Top Skills: FigmaJavaScriptReactTypescriptVue
4 Hours Ago
Easy Apply
Remote
Easy Apply
Expert/Leader
Expert/Leader
Big Data • Fintech • Mobile • Payments • Financial Services
Leads Affirm’s global customer operations vendor strategy and BPO network, owning governance, performance, compliance, resiliency, cost optimization, AI-enabled innovation, and customer experience. Manages Vendor Managers and international stakeholders while overseeing vendor selection, migrations, escalations, operational transformations, analytics, workforce planning, and recovery efforts. Builds scalable governance and high-performing, customer-centric vendor organizations across global markets.
Top Skills: Artificial Intelligence (Ai)AutomationBpo

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account