Blackpoint Cyber Logo

Blackpoint Cyber

Director of SRE

Reposted One Month Ago
Be an Early Applicant
Remote
Hiring Remotely in Canada
Expert/Leader
Remote
Hiring Remotely in Canada
Expert/Leader
Lead global SRE team to design, operate, and optimize scalable cloud infrastructure (AWS/Azure/GCP). Own reliability, observability, incident response, IaC, CI/CD, cost optimization (FinOps), security hygiene, and mentor engineers while contributing hands-on and researching AI-driven SRE tooling.
The summary above was generated by AI

Blackpoint Cyber is the leading provider of world-class cybersecurity threat hunting, detection and remediation technology. Founded by former National Security Agency (NSA) cyber operations experts who applied their learnings to bring national security-grade technology solutions to commercial customers around the world, Blackpoint Cyber is in hyper-growth mode,  fueled by a recent $190m series C round. 

Why Blackpoint?

Ready to help give some hackers hell? At Blackpoint Cyber, we fight unfair fights, eliminating threats before they strike. Built by former US Department of Defense and Intelligence security experts, our mission is to provide absolute and unified Managed Detection and Response (MDR) services to organizations worldwide 24 by 7, 365 days a year.

As a Director of SRE, you will lead the global infrastructure, reliability, and cost optimization team supporting Blackpoint Cyber’s mission-critical services. You will be responsible for the SRE function, ensuring the scalability, availability, and efficiency of our cloud infrastructure, while also supporting finops efforts to maintain cost efficiency. You will also help design the next generation of SRE at Blackpoint, helping innovate with cutting edge technologies and AI tooling to help maintain uptime and availability around the clock.

Company Culture

We value high-quality execution, ownership, and integrity—principles that are never compromised. Our team is collaborative, energetic, and thrives in a high-performance culture, continuously growing by tackling the toughest challenges in cybersecurity. This role requires strong leadership, hands-on infrastructure expertise, and a deep understanding of cost-effective scaling strategies. You will work closely with engineering, support, security, and product teams to ensure our systems are resilient, secure, and cost-efficient.

Key Responsibilities

Infrastructure & Reliability

  • Lead the design, implementation, and management of scalable, reliable, and highly available cloud-based infrastructure (AWS/Azure).

  • Establish SRE best practices, including monitoring, incident response, capacity planning, and performance tuning.

  • Improve observability, monitoring, and alerting, ensuring quick detection and resolution of reliability issues.

  • Drive automation-first approaches, reducing manual intervention through Infrastructure-as-Code (IaC) and CI/CD pipelines.

  • Lead a team of SREs, applying Blackpoint Cyber's management values of Coach, Model, Care, in defining business-critical outcomes, creating action plans, and supporting the team in achieving them.

  • Continue hands-on contributions in an SRE role

  • Design, implement, and support key infrastructure, including automated attack infrastructure deployment, isolated identity and productivity environments, and secure data storage.

  • Establish and apply security hygiene and monitoring policies to meet Blackpoint Cyber security requirements.

  • R&D with AI tooling and other SRE tools such as Grafana IRN, Loki, and Alloy

COGS Optimization & Cost Efficiency

  • Monitor and optimize cloud spending, ensuring cost-effective resource utilization without compromising reliability.

  • Define and implement cost-saving strategies (e.g., right-sizing instances, leveraging spot instances, optimizing storage, etc.).

  • Work closely with finance and procurement teams to forecast infrastructure costs and align expenses with business objectives.

Leadership & Collaboration

  • Manage and mentor a global team of SREs, DevOps engineers, and cloud infrastructure specialists.

  • Partner with engineering teams to design reliable and scalable architectures, embedding reliability into development workflows.

  • Collaborate with security teams to ensure compliance, security hardening, and disaster recovery readiness.

  • Drive post-incident reviews, ensuring continuous improvement in system resilience.

What You Bring

Must-Have Qualifications

  • 10+ years of experience in SRE, DevOps, or Cloud Infrastructure roles.

  • 5+ years of experience in people management, leading SRE team.

  • Strong experience with AWS, Azure, or GCP, with expertise in cost management and scaling strategies.

  • Proficiency in Infrastructure-as-Code (IaC) (e.g., Terraform, CloudFormation, Pulumi).

  • Hands-on experience with CI/CD pipelines, Kubernetes, and container orchestration.

  • Expertise in monitoring, logging, and observability tools (e.g., Prometheus, Grafana, Datadog, Splunk).

  • Proven ability to optimize cloud costs (COGS) while maintaining reliability and performance.

  • Strong leadership, collaboration, and problem-solving skills.

  • Experience with SLA/SLO/SLIs will be valuable.

  • A general understanding of the modern AI tooling landscape and how SRE can use that to increase velocity and improve stability

Nice-to-Have

  • Experience working in a cybersecurity or high-security environment.

  • Understanding of compliance frameworks (SOC2, ISO 27001, FedRAMP, etc.).

  • Knowledge of serverless architectures and edge computing.

  • Experience working with FinOps teams to manage cloud costs effectively.

Blackpoint Cyber welcomes and encourages applications from qualified individuals of all races, colors, religions, sex, sexual orientation, gender identity or expression, national origin, age, marital status, or any other legally protected status. We are committed to equality of opportunity in all aspects of employment.

For eligible employees in the US, Blackpoint offers competitive Health, Vision, Dental, and Life Insurance plans, a robust 401k plan, Discretionary Time Off, and other minor perks. International employees receive competitive benefits in accordance with local market standards and applicable country requirements.

Blackpoint believes all employees should share in the company’s success – equity participation is available to employees globally, with program details varying by location and employment structure.

Similar Jobs

57 Minutes Ago
Easy Apply
Remote or Hybrid
Canada
Easy Apply
Expert/Leader
Expert/Leader
Marketing Tech • Social Media • Software • Analytics • Business Intelligence
Own the strategy, roadmap, delivery, pricing, and growth of Sprout Social’s Listening product. Drive adoption and competitive differentiation through customer discovery, market analysis, product investments, and cross-functional collaboration with Engineering, Design, Marketing, Sales, Customer Experience, and GTM teams. Establish success metrics, optimize usage and pricing opportunities, communicate product direction, and lead roadmap execution across distributed teams.
Top Skills: AnalyticsProduct AnalyticsSaaSSocial Media Intelligence
3 Hours Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Set technical strategy for backend identity decisioning systems, lead major engineering initiatives, and ensure highly available operations. Partner with product, design, and analytics teams; establish coding, architecture, monitoring, testing, and on-call standards; communicate technical decisions; and mentor engineers. Build and scale distributed systems supporting KYC, user lifecycle management, and identity decisions across international markets.
Top Skills: SparkAWSKotlinKubernetesMySQLPython
3 Hours Ago
Easy Apply
Remote or Hybrid
6 Locations
Easy Apply
Junior
Junior
Big Data • Cloud • Software • Database
Build and maintain agent skills and the platform for authoring, evaluating, publishing, and monitoring them. Implement CLIs, libraries, CI integrations, evaluation harnesses, metrics, and safety/quality gates. Investigate failures, design datasets and workflows, and collaborate with cross-functional teams to improve agent behavior and automation reliability.
Top Skills: Ci/CdClisGithub ActionsMcpMongoDBStatic Analysis

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account