Netskope Logo

Netskope

Sr. DevOps Engineer, Data Platform

Posted 11 Days Ago
Remote
Hiring Remotely in India
Senior level
Remote
Hiring Remotely in India
Senior level
Operate and improve production data infrastructure across large-scale distributed cloud systems. Monitor performance, manage alerts, respond to incidents, perform root cause analysis, automate repetitive tasks, support CI/CD pipelines, document runbooks, and assist with capacity planning. The role requires participation in an on-call rotation and collaboration with Engineering to maintain reliability, efficiency, and uptime.
The summary above was generated by AI
Join the Future of Security at Netskope

Netskope (NASDAQ: NTSK) is a leader in modern security and networking for the cloud and AI era. We secure and accelerate cloud, data, and AI in real time, everywhere. Thousands of customers, including more than 30 of the Fortune 100, trust the Netskope One platform, its Zero Trust Engine, and the powerful NewEdge network to gain full visibility and control without performance trade-offs.

At Netskope, our technology is driven by our greatest strength: our people. We believe that belonging powers innovation, and success is both personal and organizational. We embrace differences in gender, ethnicity, beliefs, ability, and identity, creating an environment where every voice is heard and respected. We empower our employees to bring their authentic selves to work, grow their careers through continuous education and mentorship, and lead with transparency and curiosity. Join a team where you belong, where you are encouraged to be an entrepreneur, and where together, we continue to redefine the landscape of security.

Visit Careers at Netskope to learn more. Follow us on LinkedIn and Instagram.

About the Role:

Please note, this team is hiring across all levels and candidates are individually assessed and appropriately leveled based upon their skills and experience.

We're looking for an engineer to ensure the reliable operation of production environments for our Data Infrastructure and products, running at scale on large-volume distributed cloud systems. You'll focus on maximizing system reliability, automating routine tasks, and improving production efficiency.

This role offers hands-on exposure to modern cloud technologies — Docker, Kubernetes, networking, and platforms like AWS and GCP — while you contribute directly to system uptime and user experience for a large-scale distributed application.

You'll lead production monitoring and incident response, drive automation to reduce manual work, and collaborate with Engineering on root cause analysis, CI/CD support, and capacity planning.


What's in it for you:

In this role, you will be responsible for seamless operation of production environments for our Data Infrastructure and products within large-scale, high-volume distributed cloud systems. You will concentrate on maximizing system reliability, automating routine tasks, and maintaining the efficiency of our production systems.

This role provides practical cloud exposure to help you deepen your expertise in modern technology stacks, such as Docker, Kubernetes, Networking, and major public cloud providers like AWS and GCP. You will have the chance to contribute to enhancing user experience and system uptime for a large-scale distributed application.

What you will be doing:

  • Monitoring: Use observability dashboards to monitor system performance, error rates, and resource utilization. Define new dashboards and alerts as and when required and write technical runbooks for Incident response.
  • Incident Response: Act as the initial point of contact for production alerts, mitigate/solve the ongoing issues, escalate highly complex problems to the Engineering team, and conduct thorough root cause analysis for incidents.
  • Automation: Identify repetitive manual tasks and streamline them through automation.
  • Support CI/CD: Help create, test, and maintain automated application deployment pipelines.
  • Capacity Management: Assist in tracking system resource usage (CPU, memory, storage) and traffic volume to help forecast and scale infrastructure needs

Required skills and experience:

  • 5+ years of overall industry experience in a relevant technical role
  • Proficiency in at least one scripting language, preferably Python
  • Hands-on experience with containerization and orchestration technologies such as Docker, Kubernetes, and related cluster concepts
  • Strong understanding of public cloud infrastructure, with a preference for AWS (GCP experience also considered)
  • Working knowledge of Linux/Unix command-line environments and core web protocols including HTTP, gRPC, DNS, and TCP/IP
  • Experience with observability and monitoring tools such as Grafana and Prometheus is a plus
  • Familiarity with Infrastructure as Code (IaC) using Terraform, along with workflow automation via GitHub Actions, is a plus
  • Willingness to participate in an on-call rotation and respond to production incidents with flexibility to support critical systems

Education

  • BSCS or equivalent required, MSCS or equivalent strongly preferred

#LI-JB3


Netskope is committed to implementing equal employment opportunities for all employees and applicants for employment. Netskope does not discriminate in employment opportunities or practices based on religion, race, color, sex, marital or veteran statues, age, national origin, ancestry, physical or mental disability, medical condition, sexual orientation, gender identity/expression, genetic information, pregnancy (including childbirth, lactation and related medical conditions), or any other characteristic protected by the laws or regulations of any jurisdiction in which we operate.

Netskope respects your privacy and is committed to protecting the personal information you share with us, please refer to Netskope's Privacy Policy for more details.

The application window for this position is expected to close within 50 days. You may apply by filling out the below information, or visiting our Netskope Careers site.

Similar Jobs

3 Hours Ago
Remote or Hybrid
Senior level
Senior level
Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Lead reliability, scalability, observability, automation, infrastructure, disaster recovery, and security initiatives for Crunchyroll’s cloud-native data platforms. Establish SRE practices including SLIs, SLOs, error budgets, incident management, and postmortems. Operate Kubernetes and GCP environments, implement Infrastructure as Code, optimize capacity and performance, and drive vulnerability remediation, penetration-testing support, and cloud platform security.
Top Skills: Ci/CdDatadogGCPGoGrafanaIdentity And Access ManagementInfrastructure As CodeJavaKubernetesLinuxOpentelemetryOwasp Top 10PrometheusPythonShellTerraform
4 Hours Ago
Easy Apply
Remote or Hybrid
Easy Apply
Junior
Junior
Artificial Intelligence • Big Data • Logistics • Machine Learning • Software • Transportation
The Product Manager will oversee the Inventory & Orders module, handling product specifications, customer feedback, and integration requirements, while collaborating with cross-functional teams.
Top Skills: ConfluenceJIRAOracleSAP
7 Hours Ago
Remote
Senior level
Senior level
Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Cybersecurity • Data Privacy
Manages global partner program operations, including case resolution, workflow optimization, partner onboarding, content governance, certification reporting, AI-driven analytics, and financial processes. The role coordinates across Legal, Sales, Finance, and Product, uses Salesforce and Coupa, and delivers data-backed recommendations to leadership. It requires strong problem-solving, communication, adaptability, and experience supporting global channel or partner operations.
Top Skills: ClaudeCoupaEnterprise Search PlatformsGeminiGenerative AiGleanGoogle WorkspaceSalesforce

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account