Moxie (joinmoxie.com) Logo

Moxie (joinmoxie.com)

Staff Platform Engineer (IC-4)

Reposted 19 Days Ago
In-Office or Remote
Hiring Remotely in Australia
Senior level
In-Office or Remote
Hiring Remotely in Australia
Senior level
Ownership of platform reliability and incident response across cloud infrastructure. Improve observability, CI/CD pipelines, deployments, developer experience, and operational tooling. Act as primary on-call platform expert, collaborate with product teams, and run postmortems to raise platform reliability.
The summary above was generated by AI

At Moxie, we empower ambitious aesthetic entrepreneurs to build profitable, independent practices—without burnout, overwhelm, or guesswork. In just a few years, we've grown from an idea to a global, remote-first team now supporting 700+ practices nationwide.

Our purpose is simple: to unlock sustainable success for aesthetic entrepreneurs, at every stage of their journey.

Staff Platform Engineer

Remote, Full-time
Location:
LATIN AMERICA Colombia, Brazil, Dominican Republic, or Chile - Fully Remote (Work from Home)
Working hours: Core overlap with 9 AM – 5 PM EST (flexible schedules between 7 AM – 8 PM EST)

About Moxie

Moxie empowers aesthetic industry professionals to become successful entrepreneurs. We provide a sophisticated SaaS platform that simplifies the operational complexities of running MedSpas, enabling nurses and medical professionals to launch, operate, and grow their businesses across the country.

Hundreds of customers rely on Moxie Suite to run their MedSpas end-to-end: scheduling, medical purchasing, payments and invoicing, bookkeeping, analytics, and more. The platform operates at real-world scale and reliability requirements, integrating with systems such as AWS, Vercel, Cloudflare, Stripe, Twilio, Datadog, and others.

We are a fast-growing company focused on building reliable infrastructure, strong developer experience, and operational excellence as we scale.

 
 
The Role

We are looking for a Staff Platform Engineer to help own and evolve the systems that enable Moxie engineers to ship safely, quickly, and reliably.

This role sits at the intersection of DevOps and SRE. Your most important responsibility will be incident handling — you'll be the primary owner of detecting, responding to, and resolving production issues across our infrastructure. During US business hours, you'll typically be the sole platform/infra expert on call, though developers will be available to support you as needed. Given this, the role can be demanding at times and requires availability beyond standard hours when incidents arise.

Beyond incident response, you'll work on cloud infrastructure, CI/CD pipelines, deployment workflows, local development environments, observability, and operational tooling. You'll partner closely with product engineering teams, but your primary responsibility is the health, reliability, and usability of the platform itself.

This is a hands-on individual contributor role with meaningful ownership, but no people management responsibilities.

Key ResponsibilitiesObservability & Reliability
  • Participate in incident response as needed

  • Own and improve monitoring, logging, and alerting using Datadog, AWS, Vercel and related tools

  • Ensure systems are observable and failure modes are well understood

  • Help teams learn from incidents through postmortems and follow-ups

  • Balance reliability with delivery speed through pragmatic SRE practices

Platform & Infrastructure
  • Own and operate core platform systems across AWS, GCP, Vercel, Github, and Cloudflare

  • Improve reliability, scalability, and security of production and non-production environments

  • Maintain and evolve infrastructure supporting multiple services and teams

CI/CD & Deployments
  • Own and improve CI/CD pipelines (GitHub Actions), focusing on speed, reliability, and clarity

  • Improve deployment workflows, rollbacks, and environment consistency

  • Reduce deployment-related risk and manual intervention

  • Partner with engineers and our QA team to improve release confidence and velocity

Developer Experience & Local Development
  • Improve local development environments and onboarding experience for engineers

  • Reduce friction in common workflows (setup, testing, debugging)

  • Maintain tooling and documentation that helps engineers move faster with confidence

Cross-Team Collaboration
  • Work closely with the product engineering team to understand platform pain points and improve local development experience.

  • Provide guidance and support on infrastructure, deployments, and operational best practices

  • Contribute to platform standards and shared tooling through collaboration, not mandates

QualificationsRequired
  • 5+ years of experience in platform, DevOps, or SRE-focused roles

  • Strong experience operating production systems on AWS (certifications strongly preferred), Vercel, and/or GCP

  • Experience building and maintaining CI/CD pipelines (GitHub Actions or similar)

  • Strong understanding of cloud networking, security fundamentals, and IAM

  • Experience with observability tooling (Datadog preferred)
    Ability to troubleshoot production issues calmly and systematically

  • Excellent written and verbal communication skills in English (C1 or higher).

Nice to Have
  • Experience with Cloudflare (DNS, WAF, edge configuration)

  • Experience with Agentic and LLM based tooling and automations (Cursor, Codex, Claude Code, etc)

  • Experience improving local development tooling (ie Docker, Husky, bash)

  • Familiarity with infrastructure-as-code (Terraform or similar)

  • Experience supporting regulated or compliance-sensitive environment

  • Experience working with PII, PHI, and sensitive data systems in general.

Our Stack
  • Cloud & Infrastructure: AWS ECS/Fargate, GCP, and Vercel

  • Edge & Security: Cloudflare

  • CI/CD: GitHub Actions

  • Observability: Datadog

  • Database: AWS RDS (postgres)

  • Version Control: Git, GitHub

  • AI / LLM Tooling: Claude Code, Gemini, Cursor, CodeRabbit, Glean, Codex

  • Backend Services: Python, Django (operational ownership, not feature dev)

At Moxie, we believe in creating a workplace where everyone feels valued, trusted, and included. Our team lives by our values: act as owners, give more than we take, move with speed and care, and simplify and learn every day.

We welcome people of all backgrounds, experiences, and perspectives to apply. If you require any accommodations to fully participate in the interview process, please let us know, we’re happy to assist.

Similar Jobs

Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Create high-quality 3D character gameplay animations for navigation, interactions, conversations, traversal, and other gameplay systems. Export, implement, and test animations in the game engine; troubleshoot animation issues; collaborate with animators, designers, and gameplay programmers; participate in reviews; and mentor junior animators. The role requires strong knowledge of body mechanics, biped locomotion, motion capture, keyframe animation, gameplay systems, and animation workflows, with experience shipping an AAA game.
Top Skills: Animation Blend SystemsAnimation GraphsKeyframe AnimationMayaMotion CaptureMotion MatchingMotionbuilderRiggingRoot MotionUnreal Engine
11 Hours Ago
Remote
Internship
Internship
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Participate in a 15-week full-time software engineering internship in Vancouver, Burnaby, or Richmond. Build and ship product features, contribute code to Atlassian products, apply data structures and algorithms, and learn full-lifecycle development through mentorship and collaboration with senior engineers. Work may involve cloud initiatives and AI-driven features.
Top Skills: CC++JavaPython
13 Hours Ago
Easy Apply
Remote
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Own and evolve GitLab’s authorization systems across its Ruby on Rails monolith and next-generation Rust policy engine. Design fine-grained permissions for users, tokens, roles, and AI agents; secure GraphQL and REST APIs; lead feature-flagged rollouts and migrations; improve performance and reliability; and collaborate across authentication, platform, AI, and modular-service teams in a distributed asynchronous environment.
Top Skills: CedarGoGraphQLGrpcProtocol BuffersRestRubyRuby On RailsRustYamlZanzibar-Style Authorization

What you need to know about the Sydney Tech Scene

From opera to comedy shows, the Sydney Opera House hosts more than 1,600 performances a year, yet its entertainment sector isn't the only one taking center stage. The city's tech sector has earned a reputation as one of the fastest-growing in the region. More specifically, its IT sector stands out as the country's third-largest, growing at twice the rate of overall employment in the past decade as businesses continue to digitize their operations to stay competitive.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account