Staff Site Reliability Engineer - Network Infrastructure
Quick Summary
Global Ingress & Routerless Evolution (CFSaaS): Design and engineer Next-Gen traffic entryways utilizing Cloudflare for SaaS. Drive the migration toward modern,
8+ years of experience in production Cloud Infrastructure Engineering, Systems Engineering, or Core Cloud Networking roles.
Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities, such as reducing costs and doing more for your customers.
If you like to be challenged and have a passion for solving large-scale automation, global traffic routing, and resilient cloud infrastructure problems, we would love to hear from you. The ideal candidate is someone who exemplifies the ethics of, “If you have to do something more than once, automate it” and who can rapidly self-educate on new network topologies, multi-cloud ecosystems, and agentic engineering models.
The Staff Software Engineer - Infrastructure will play a foundational role in architecting, evolving, and securing Okta's global core network fabric and multi-cloud platform layers. This position focuses on building highly resilient, edge management in AWS and GCP cloud, executing enterprise-wide cloud migrations (AWS to GCP), enforcing absolute Zero-Trust primitives (mTLS / TLS 1.3), and integrating cutting-edge AI automation layers (Agentic SRE) into our day-to-day operations.
As a Staff Engineer on this team, you will act as a key technical anchor in India, working closely with global counterparts to maintain Okta's high-availability SLAs while protecting the platform against active global DDoS attacks.
Responsibilities
~2 min read- →Global Ingress & Routerless Evolution (CFSaaS): Design and engineer Next-Gen traffic entryways utilizing Cloudflare for SaaS. Drive the migration toward modern, decentralized edge infrastructure—eliminating legacy hardware routing bottlenecks, shifting compute to Edge Workers, and lowering latency for millions of global requests.
- →Multi-Cloud & Core Networking: Build, scale, and maintain Okta's multi-cloud backbone across AWS and GCP. Actively architect clean hub-and-spoke models utilizing AWS Transit Gateway (TGW), AWS Global Accelerator (AGA), and GCP Network Connectivity Center to unify converged enterprise product networks (Harmony).
- →Zero-Trust Cryptographic Enforcement: Implement and manage strict cryptographic security controls at scale, driving the enforcement of TLS 1.3 externally and deep mTLS (Mutual TLS) mesh architectures internally across microservice proxies (Envoy/Nginx).
- →Agentic SRE & AI Automation: Innovate on platform operations by designing deterministic playbooks and automated telemetry ingestion pipelines using AI models (LLMs/Agents). Move the team from reactive debugging to autonomous anomaly detection and self-healing systems.
- →DDoS Firefight Ownership & Mitigation: Provide first-line technical triage during high-volume volumetric and application-layer (L7) DDoS attacks during India Data Center (IDC) hours. Analyze real-time proxy/VPC flow logs and implement rapid WAF mitigations to preserve cell isolation and tenant availability.
- →Platform Automation (GitOps): Champion Infrastructure-as-Code (IaC) principles. Continuously identify and eliminate network configuration drifts across heterogeneous cloud landing zones by maintaining high-quality Terraform and automation modules.
- →Documentation & Cross-Border Mentorship: Create detailed global architecture blueprints, technical documentation, and disaster recovery runbooks. Act as a technical mentor to a lean, high-performing 3-member local team while collaborating smoothly with US-time engineering leaders.
Requirements
~2 min read- 8+ years of experience in production Cloud Infrastructure Engineering, Systems Engineering, or Core Cloud Networking roles.
- 5+ years of hands-on experience with deep enterprise cloud networking constructs: AWS Transit Gateway (TGW), VPC peering, AWS Global Accelerator, Direct Connect, and GCP networking equivalents (Shared VPCs, Cloud Routers).
- Strong working knowledge of Edge Delivery networks and edge compute stacks (Cloudflare for SaaS, Cloudflare Tunnels, Custom Hostname routing, Edge Workers).
- Deep understanding of network layers and security protocols: HTTP/S routing, TCP/IP stack debugging (tcpdump, Wireshark), Public Key Infrastructure (PKI), certificate rotation pipelines, and proxy layers (Envoy, Nginx, API Gateways).
- Advanced proficiency with Infrastructure-as-Code (IaC) tools, specifically Terraform, for managing complex, multi-account and multi-region landing zones.
- Robust scripting and software development skills in Python or Go for building infrastructure tools, log parsers, and platform automation frameworks.
- Expertise with large-scale telemetry aggregation, logging, and monitoring systems (Splunk, Datadog, Prometheus, Grafana, AWS CloudWatch, or VPC Flow Logs).
- Familiarity with containerized infrastructure elements (Docker, basic Kubernetes deployments) as they interface with public network load balancers.
- Practical experience or strong architectural interest in implementing AI-assisted automation layers (LLM APIs, agentic orchestration frameworks) for parsing infrastructure metrics or automating repetitive on-call tasks.
- Prior experience supporting high-availability, multi-tenant SaaS platforms with strict cell-based architectures and zero-downtime requirements.
- Bachelor’s degree in Computer Science, Computer Engineering, or a related technical field (or equivalent professional experience).
- Certifications (Preferred): AWS Certified Advanced Networking - Specialty, Google Cloud Certified Professional Cloud Network Engineer, or equivalent industry networking credentials.
#LI_Hybrid
P24489_3505296
- Supporting Your Well-Being
- Driving Social Impact
- Developing Talent and Fostering Connection + Community
We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.
Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.
If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.
Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
Location & Eligibility
Listing Details
- Posted
- July 28, 2026
- First seen
- July 28, 2026
- Last seen
- July 28, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 67%
- Scored at
- July 28, 2026
Signal breakdown

The foundation for secure connections between people and technology.
View company profilePlease let Okta know you found this job on Jobera.
Similar Staff Site Reliability Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.