Moxie (joinmoxie.com) Logo

Moxie (joinmoxie.com)

Staff Platform Engineer (LATAM)

Reposted One Month Ago
In-Office or Remote
Hiring Remotely in Canada
Senior level
In-Office or Remote
Hiring Remotely in Canada
Senior level
Ownership of platform reliability and incident response across cloud infrastructure. Improve observability, CI/CD pipelines, deployments, developer experience, and operational tooling. Act as primary on-call platform expert, collaborate with product teams, and run postmortems to raise platform reliability.
The summary above was generated by AI

At Moxie, we empower ambitious aesthetic entrepreneurs to build profitable, independent practices—without burnout, overwhelm, or guesswork. In just a few years, we've grown from an idea to a global, remote-first team now supporting 700+ practices nationwide.

Our purpose is simple: to unlock sustainable success for aesthetic entrepreneurs, at every stage of their journey.

Staff Platform Engineer

Remote, Full-time
Location:
LATIN AMERICA Colombia, Brazil, Dominican Republic, or Chile - Fully Remote (Work from Home)
Working hours: Core overlap with 9 AM – 5 PM EST (flexible schedules between 7 AM – 8 PM EST)

About Moxie

Moxie empowers aesthetic industry professionals to become successful entrepreneurs. We provide a sophisticated SaaS platform that simplifies the operational complexities of running MedSpas, enabling nurses and medical professionals to launch, operate, and grow their businesses across the country.

Hundreds of customers rely on Moxie Suite to run their MedSpas end-to-end: scheduling, medical purchasing, payments and invoicing, bookkeeping, analytics, and more. The platform operates at real-world scale and reliability requirements, integrating with systems such as AWS, Vercel, Cloudflare, Stripe, Twilio, Datadog, and others.

We are a fast-growing company focused on building reliable infrastructure, strong developer experience, and operational excellence as we scale.

 
 
The Role

We are looking for a Staff Platform Engineer to help own and evolve the systems that enable Moxie engineers to ship safely, quickly, and reliably.

This role sits at the intersection of DevOps and SRE. Your most important responsibility will be incident handling — you'll be the primary owner of detecting, responding to, and resolving production issues across our infrastructure. During US business hours, you'll typically be the sole platform/infra expert on call, though developers will be available to support you as needed. Given this, the role can be demanding at times and requires availability beyond standard hours when incidents arise.

Beyond incident response, you'll work on cloud infrastructure, CI/CD pipelines, deployment workflows, local development environments, observability, and operational tooling. You'll partner closely with product engineering teams, but your primary responsibility is the health, reliability, and usability of the platform itself.

This is a hands-on individual contributor role with meaningful ownership, but no people management responsibilities.

Key ResponsibilitiesObservability & Reliability
  • Participate in incident response as needed

  • Own and improve monitoring, logging, and alerting using Datadog, AWS, Vercel and related tools

  • Ensure systems are observable and failure modes are well understood

  • Help teams learn from incidents through postmortems and follow-ups

  • Balance reliability with delivery speed through pragmatic SRE practices

Platform & Infrastructure
  • Own and operate core platform systems across AWS, GCP, Vercel, Github, and Cloudflare

  • Improve reliability, scalability, and security of production and non-production environments

  • Maintain and evolve infrastructure supporting multiple services and teams

CI/CD & Deployments
  • Own and improve CI/CD pipelines (GitHub Actions), focusing on speed, reliability, and clarity

  • Improve deployment workflows, rollbacks, and environment consistency

  • Reduce deployment-related risk and manual intervention

  • Partner with engineers and our QA team to improve release confidence and velocity

Developer Experience & Local Development
  • Improve local development environments and onboarding experience for engineers

  • Reduce friction in common workflows (setup, testing, debugging)

  • Maintain tooling and documentation that helps engineers move faster with confidence

Cross-Team Collaboration
  • Work closely with the product engineering team to understand platform pain points and improve local development experience.

  • Provide guidance and support on infrastructure, deployments, and operational best practices

  • Contribute to platform standards and shared tooling through collaboration, not mandates

QualificationsRequired
  • 5+ years of experience in platform, DevOps, or SRE-focused roles

  • Strong experience operating production systems on AWS (certifications strongly preferred), Vercel, and/or GCP

  • Experience building and maintaining CI/CD pipelines (GitHub Actions or similar)

  • Strong understanding of cloud networking, security fundamentals, and IAM

  • Experience with observability tooling (Datadog preferred)
    Ability to troubleshoot production issues calmly and systematically

  • Excellent written and verbal communication skills in English (C1 or higher).

Nice to Have
  • Experience with Cloudflare (DNS, WAF, edge configuration)

  • Experience with Agentic and LLM based tooling and automations (Cursor, Codex, Claude Code, etc)

  • Experience improving local development tooling (ie Docker, Husky, bash)

  • Familiarity with infrastructure-as-code (Terraform or similar)

  • Experience supporting regulated or compliance-sensitive environment

  • Experience working with PII, PHI, and sensitive data systems in general.

Our Stack
  • Cloud & Infrastructure: AWS ECS/Fargate, GCP, and Vercel

  • Edge & Security: Cloudflare

  • CI/CD: GitHub Actions

  • Observability: Datadog

  • Database: AWS RDS (postgres)

  • Version Control: Git, GitHub

  • AI / LLM Tooling: Claude Code, Gemini, Cursor, CodeRabbit, Glean, Codex

  • Backend Services: Python, Django (operational ownership, not feature dev)

At Moxie, we believe in creating a workplace where everyone feels valued, trusted, and included. Our team lives by our values: act as owners, give more than we take, move with speed and care, and simplify and learn every day.

We welcome people of all backgrounds, experiences, and perspectives to apply. If you require any accommodations to fully participate in the interview process, please let us know, we’re happy to assist.

Similar Jobs

Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Leads production delivery for a large-scale AAA game pillar. Translates priorities into sprint plans, facilitates Scrum ceremonies, manages backlogs, risks, dependencies, milestones, and delivery visibility, and removes blockers across multiple multidisciplinary pods. Partners with Product Owners, pod leads, and craft directors to improve workflows, team health, communication, and accountability. Requires extensive game production experience, agile facilitation, and cross-team coordination.
Top Skills: JIRAKanbanScrum
Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Develop reinforcement learning systems for autonomous agents by designing high-fidelity 2D/3D simulation environments, reward functions, and policy architectures. Implement and optimize algorithms such as PPO, SAC, and Offline RL for high-dimensional observations. Develop sim-to-real strategies using domain randomization and adaptation, collaborate with ML, annotation, and program management teams, and ensure reliable agent performance and environment parity.
Top Skills: BulletCleanrlConfluenceExperiment Tracking FrameworksGitGit ServerIsaac SimJIRAMujocoPythonRay RllibSlackStable Baselines3UnityUnix ShellUnreal Engine
An Hour Ago
Easy Apply
Remote
Canada
Easy Apply
Mid level
Mid level
Artificial Intelligence • Fintech • Hardware • Information Technology • Sales • Software • Transportation
Manage business operations for the Services organization, including Implementation, Installation, and Customer Education. Optimize workflows, monitor performance, manage recurring business processes, produce Salesforce reports, and lead strategic projects that improve customer experience and team efficiency. Partner with Sales, Support, Product, and leadership on forecasting, territory alignment, QBRs, resource allocation, and operational decision-making in a high-growth SaaS environment.
Top Skills: Ai-Driven ToolsSalesforceSFDCSpreadsheets

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account