Netskope Logo

Netskope

Sr. DevOps Engineer, Data Platform

Posted 8 Days Ago
Remote
Hiring Remotely in India
Senior level
Remote
Hiring Remotely in India
Senior level
Operate and improve production data infrastructure across large-scale distributed cloud systems. Monitor performance, manage alerts, respond to incidents, perform root cause analysis, automate repetitive tasks, support CI/CD pipelines, document runbooks, and assist with capacity planning. The role requires participation in an on-call rotation and collaboration with Engineering to maintain reliability, efficiency, and uptime.
The summary above was generated by AI
Join the Future of Security at Netskope

Netskope (NASDAQ: NTSK) is a leader in modern security and networking for the cloud and AI era. We secure and accelerate cloud, data, and AI in real time, everywhere. Thousands of customers, including more than 30 of the Fortune 100, trust the Netskope One platform, its Zero Trust Engine, and the powerful NewEdge network to gain full visibility and control without performance trade-offs.

At Netskope, our technology is driven by our greatest strength: our people. We believe that belonging powers innovation, and success is both personal and organizational. We embrace differences in gender, ethnicity, beliefs, ability, and identity, creating an environment where every voice is heard and respected. We empower our employees to bring their authentic selves to work, grow their careers through continuous education and mentorship, and lead with transparency and curiosity. Join a team where you belong, where you are encouraged to be an entrepreneur, and where together, we continue to redefine the landscape of security.

Visit Careers at Netskope to learn more. Follow us on LinkedIn and Instagram.

About the Role:

Please note, this team is hiring across all levels and candidates are individually assessed and appropriately leveled based upon their skills and experience.

We're looking for an engineer to ensure the reliable operation of production environments for our Data Infrastructure and products, running at scale on large-volume distributed cloud systems. You'll focus on maximizing system reliability, automating routine tasks, and improving production efficiency.

This role offers hands-on exposure to modern cloud technologies — Docker, Kubernetes, networking, and platforms like AWS and GCP — while you contribute directly to system uptime and user experience for a large-scale distributed application.

You'll lead production monitoring and incident response, drive automation to reduce manual work, and collaborate with Engineering on root cause analysis, CI/CD support, and capacity planning.


What's in it for you:

In this role, you will be responsible for seamless operation of production environments for our Data Infrastructure and products within large-scale, high-volume distributed cloud systems. You will concentrate on maximizing system reliability, automating routine tasks, and maintaining the efficiency of our production systems.

This role provides practical cloud exposure to help you deepen your expertise in modern technology stacks, such as Docker, Kubernetes, Networking, and major public cloud providers like AWS and GCP. You will have the chance to contribute to enhancing user experience and system uptime for a large-scale distributed application.

What you will be doing:

  • Monitoring: Use observability dashboards to monitor system performance, error rates, and resource utilization. Define new dashboards and alerts as and when required and write technical runbooks for Incident response.
  • Incident Response: Act as the initial point of contact for production alerts, mitigate/solve the ongoing issues, escalate highly complex problems to the Engineering team, and conduct thorough root cause analysis for incidents.
  • Automation: Identify repetitive manual tasks and streamline them through automation.
  • Support CI/CD: Help create, test, and maintain automated application deployment pipelines.
  • Capacity Management: Assist in tracking system resource usage (CPU, memory, storage) and traffic volume to help forecast and scale infrastructure needs

Required skills and experience:

  • 5+ years of overall industry experience in a relevant technical role
  • Proficiency in at least one scripting language, preferably Python
  • Hands-on experience with containerization and orchestration technologies such as Docker, Kubernetes, and related cluster concepts
  • Strong understanding of public cloud infrastructure, with a preference for AWS (GCP experience also considered)
  • Working knowledge of Linux/Unix command-line environments and core web protocols including HTTP, gRPC, DNS, and TCP/IP
  • Experience with observability and monitoring tools such as Grafana and Prometheus is a plus
  • Familiarity with Infrastructure as Code (IaC) using Terraform, along with workflow automation via GitHub Actions, is a plus
  • Willingness to participate in an on-call rotation and respond to production incidents with flexibility to support critical systems

Education

  • BSCS or equivalent required, MSCS or equivalent strongly preferred

#LI-JB3


Netskope is committed to implementing equal employment opportunities for all employees and applicants for employment. Netskope does not discriminate in employment opportunities or practices based on religion, race, color, sex, marital or veteran statues, age, national origin, ancestry, physical or mental disability, medical condition, sexual orientation, gender identity/expression, genetic information, pregnancy (including childbirth, lactation and related medical conditions), or any other characteristic protected by the laws or regulations of any jurisdiction in which we operate.

Netskope respects your privacy and is committed to protecting the personal information you share with us, please refer to Netskope's Privacy Policy for more details.

The application window for this position is expected to close within 50 days. You may apply by filling out the below information, or visiting our Netskope Careers site.

Similar Jobs

Yesterday
Remote
Senior level
Senior level
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Design, review, and optimize high-reliability power distribution and substation systems for semiconductor fabs. Produce SLDs, sizing calculations, system studies, grounding/earthing, protection coordination, and commissioning support. Integrate SCADA/BMS, ensure compliance with standards, manage vendors/contractors, and support procurement, FAT/SAT, and energization.
Top Skills: AisArc Flash AssessmentAutocadAutocad ElectricalBimBmsChatgptClaudeCo-PilotEtapFmcsGeneratorsGisHarmonic FiltersLlmsProtection RelaysRevitRmuScadaSkm Power ToolsSwitchgearTransformersUpsVfd
Yesterday
Easy Apply
In-Office or Remote
Easy Apply
Senior level
Senior level
Cloud • Information Technology • Security • Software
Provide Tier 2 technical support for complex Windows and Active Directory integration issues in a SaaS environment. Troubleshoot escalations using logs, network traces, memory dumps, and structured testing; communicate technical causes and resolutions to enterprise customers and executives. Collaborate with engineering, product, account management, and customer success teams, mentor peers, maintain knowledge-base content, improve diagnostic workflows, and participate in an on-call rotation.
Top Skills: Active DirectoryAPIsAWSAzure Active DirectoryCredential ManagerGoogle SuiteLdapOktaPktmonPowershellRadiusScimSlackSsoSysinternalsWindowsWiresharkXpath
Yesterday
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Owns business administration and continuous improvement of Veeva Vault eQMS within a regulated pharmaceutical quality organization. Responsibilities include access management, release coordination, reporting, documentation, configuration advisory, enhancement prioritization, data operations, stakeholder communications, and cross-functional program execution. The role also evaluates AI, automation, and digital transformation opportunities while ensuring GxP compliance, data integrity, system continuity, and effective quality operations.
Top Skills: DataikuMicrosoft Power AutomatePower BISnowflakeSpotfireSQLVeeva Vault Eqms

What you need to know about the Ottawa Tech Scene

The capital city of Canada and the nation's fourth-largest urban area, Ottawa has proven a rapidly growing global tech hub. With over 1,800 tech companies, many of which are leaders in their sectors, the city's tech talent now makes up more than 13 percent of its total workforce. This growth is driven not only by the big players like UL Solutions and Dropbox, but also by a thriving startup ecosystem, as new businesses emerge to follow in the footsteps of those that came before them.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account