Site Reliability Engineer
About this role
Join the Future of Security at Netskope
Netskope (NASDAQ: NTSK) is a leader in modern security and networking for the cloud and AI era. We secure and accelerate cloud, data, and AI in real time, everywhere. Thousands of customers, including more than 30 of the Fortune 100, trust the Netskope One platform, its Zero Trust Engine, and the powerful NewEdge network to gain full visibility and control without performance trade-offs.
At Netskope, our technology is driven by our greatest strength: our people. We believe that belonging powers innovation, and success is both personal and organizational. We embrace differences in gender, ethnicity, beliefs, ability, and identity, creating an environment where every voice is heard and respected. We empower our employees to bring their authentic selves to work, grow their careers through continuous education and mentorship, and lead with transparency and curiosity. Join a team where you belong, where you are encouraged to be an entrepreneur, and where together, we continue to redefine the landscape of security.
Visit Careers at Netskope to learn more. Follow us on LinkedIn and Instagram.
About the role
As a Site Reliability Engineer you will join a team of software engineers focused on improving availability, latency, performance, efficiency, change management, observability, emergency response, and capacity planning of the Netskope platform. You will design, develop and implement strategies and solutions that improve the reliability of Netskope’s production services. You will contribute to key initiatives and innovations through designing, building, and automating complex computing systems based primarily on microservice architectures.
Essential Functions:
- Embrace our corporate culture by fostering collaboration without boundaries, encouraging teamwork and communication that is clear, open, and honest.
- Partner closely with our development teams and product managers to architect and build features that are highly available, performant and secure.
- Develop innovative ways to smartly measure, monitor & report on service and infrastructure health.
- Drive efficiencies in systems and processes: capacity planning, configuration management, performance tuning, monitoring and root cause analysis.
- Design, develop, test, document, and implement strategies, processes, automation, and solutions for Netskope’s production services.
- Understand, and be able to communicate, the nonfunctional requirements for the system and services (security, reliability, scalability, maintainability, etc.).
- Become a subject matter expert on all aspects of designated Netskope products, services, and their dependent infrastructure in order to support and improve the reliability, function, and scalability of these solutions.
Required Experience:
- Bachelor’s Degree or higher in Computer Science, Engineering, or combination of comparable education and experience typically obtained by 5 or more years related work experience.
- 3+ years experience with designing, building, and managing complex computing systems including 1-2 years experience as a Site Reliability Engineer.
- Experience with working with private or public cloud services in a distributed, highly available, and large scale production environment.
- Experience with analyzing systems and software to drive the improvement of availability, performance, scale, and efficiency of microservices.
- Experience with algorithms, data structures, complexity analysis, and software design.
- Experience with reviewing and supplying non-functional business requirements and functional specification.
- Demonstrated ability to debug and optimize code and automate routine tasks.
- Demonstrated ability and willingness to act as subject matter expert, tracking technology/industry trends, and to provide data driven reasoning for recommending technology paths.
- Excellent verbal and written communication skills.
Desired Experience:
- Demonstrated ability to sanitize and automate a rapidly growing global system.
- Extensive TCP/IP knowledge and experience with network flow analysis and tuning.
- Experience with operational support systems, automation, IaC tools, and CI/CD tools.
- Knowledge of the Agile project management methodologies.
- Knowledge in a variety of programming languages including Python, C, C++, Go, Rust.
- Experience with modern cloud and virtualization technologies (Docker, Kubernetes, AWS, GCP, KVM, OpenNebula, OpenStack or any orchestration platforms)
- Established track record for providing valuable technical details to developers and engineers.
- Established track record for providing valuable feedback during the planning and implementation of existing projects.
#LI-SC3
Netskope is committed to implementing equal employment opportunities for all employees and applicants for employment. Netskope does not discriminate in employment opportunities or practices based on religion, race, color, sex, marital or veteran statues, age, national origin, ancestry, physical or mental disability, medical condition, sexual orientation, gender identity/expression, genetic information, pregnancy (including childbirth, lactation and related medical conditions), or any other characteristic protected by the laws or regulations of any jurisdiction in which we operate.
Netskope respects your privacy and is committed to protecting the personal information you share with us, please refer to Netskope's Privacy Policy for more details.
The application window for this position is expected to close within 50 days. You may apply by filling out the below information, or visiting our Netskope Careers site.
