
Okta · Washington
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure t...
Secure Every Identity, from AI to Human
Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables
organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world
stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.
This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.
At Okta, our motto is "Always On." Within the Technical Operations (TechOps) team, we live this mission by building the most
reliable and performant systems on the planet. We empower organizations to do their most significant work by securely connecting
any person, on any device, to the technologies they need.
We are looking for an experienced Senior Site Reliability Engineer (SRE) who thrives on the challenge of managing large-scale
cloud production systems. The ideal candidate is a self-starter who lives by the ethic: "If you have to do it twice, automate it."
Based in the Washington, D.C. area, with on-site customer travel, you will ensure our infrastructure maintains uncompromising
reliability and performance while supporting the most sensitive national security missions.
Security Requirement: Must be able to obtain and maintain a U.S. security clearance (Secret or Top Secret) to the extent required
by U.S. Government contracts.
The selected candidate may be subject to drug testing to the extent required by U.S. Government contracts.
reliability.
implementing permanent preventive solutions.
technical workflows.
delivery.
debugging of Helm values and charts.
CloudFormation).
environments.
Networking: Solid understanding of networking concepts and IP protocols; experience with multi-cloud environments is a significant
plus.
#LI-Hybrid
Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois,
New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work
location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance,
401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and
policies. To learn more about our Total Rewards program please visit: https://rewards.okta.com/us.
The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado,
The Okta Experience
We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate.
Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our
mission and team from day one.
Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race,
color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental
disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions
records, consistent with applicable laws.
If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use
this Form to request an accommodation.
Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York
City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment
and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City,
please click here to view our full NYC AEDT Notice.
Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. SENIOR MANAGER, SITE RELIABILITY ENGINEERING Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. THE FEDERAL OPERATIONS ENGINEERING GROUP Okta's Federal Operations team supports government customers operating in FedRAMP-authorized, IL4, and IL5 environments. We deliver the same 99.999% availability promise to the federal market while meeting the strict security, compliance, and operational requirements that come with it. We're looking for a technical leader who understands both the SRE discipline and the unique demands of federal customer relationships — someone who can hold a technical conversation with an agency ISSO in the morning and unblock an incident bridge with an engineering team in the afternoon. As Senior Manager of Federal SRE Operations, you will own the operational health of Okta's federal environments, lead a team of engineers working within compliance-governed change processes, and serve as a trusted point of contact for federal security stakeholders. What you'll be doing * Lead and grow a team of SREs operating Okta's FedRAMP High, IL4, and IL5 environments, ensuring reliability and compliance obligations are met simultaneously. * Serve as the primary operational relationship manager for federal agency security leaders — including Authorizing Officials (AOs), ISSOs, and ISSMs — translating technical SRE practices into language and posture appropriate for federal oversight. * Own the continuous monitoring program for Okta's federal boundary: coordinate POA&M remediation, drive vulnerability SLA compliance, and represent operational risk posture in ATO reviews. * Partner with Okta's GovCloud product and security teams to ensure infrastructure changes go through FedRAMP-compliant change management processes without sacrificing velocity. * Drive incident response and post-incident review processes that satisfy federal reporting obligations (e.g., FISMA incident notification timelines) while improving system reliability. * Manage cross-functional relationships with 3PAOs, federal agency security teams, and internal compliance, legal, and security stakeholders during audits, ATO renewals, and security assessments. * Maintain operational runbooks, system security plans (SSPs), and evidence packages that reflect the true state of the environment — not just what audit season requires. * Accelerate the federal engineering team's velocity by building compliant self-service tooling, automation, and CI/CD pipelines that reduce the friction of operating in a restricted environment. * Lead, mentor, and grow engineers across the federal SRE organization, including setting career paths that bridge deep federal compliance knowledge with modern SRE practices. What you'll bring to the role * 5+ years of experience in technical leadership and people management, with at least 3 years in a federal or regulated environment (FedRAMP, DoD, IC, or equivalent). * Demonstrated experience managing relationships with federal security stakeholders — AOs, ISSOs, agency CISOs, or equivalent — not just writing documentation for them. * Working knowledge of NIST SP 800-53, FedRAMP authorization frameworks, FISMA reporting requirements, and DISA STIGs. * Experience operating large-scale cloud infrastructure in FedRAMP-authorized environments (AWS GovCloud preferred); IL4/IL5 experience a strong plus. * Strong background in SRE or platform engineering fundamentals: Kubernetes, IaC (Terraform), CI/CD, observability, and incident management. * Ability to hold compliance and reliability in tension — understanding when a change needs to go through a CAB and when a P1 incident requires you to move faster and document after. * Strong verbal and written communication skills; comfortable presenting operational risk and reliability posture to non-technical federal security officials. * Experience supporting or leading FedRAMP ATO processes, continuous monitoring programs, or FISMA annual assessments. * U.S. Citizenship required (see below). * Active or current eligibility for a federal security clearance (Secret or above) is a plus and may be required for certain program support. Additional requirements: * This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire. (P11672_3441312) #LI-Hybrid Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: https://rewards.okta.com/us. The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between: $207,000—$284,900 USD The Okta Experience * Supporting Your Well-Being * Driving Social Impact * Developing Talent and Fostering Connection + Community We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one. Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws. If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation. Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
Join Truecaller – The place where innovation meets impact! Truecaller's mission is to build trust in communication by making it safer, smarter, and more efficient. Born in Sweden, trusted by the world, and here’s why we stand out: * We are trusted by over 450 million active users every month across 190+ countries * We identify over 15 billion calls daily, helping users avoid spam and scams * We are powered by a team of 450+ employees from 45+ nationalities We always look for people who take initiative, own their work, and keep raising the bar. An entrepreneurial mindset matters here, especially when it turns bold ideas into real actions. We stay collaborative and focused, always searching for smarter paths forward. If you want to make an impact and grow with a team that inspires millions, you’ll fit right in. The role: You will design, manage and maintain infrastructure services over the course of their lifecycle. You will be responsible for building and improving the performance, reliability, availability, security, and evolution of the Truecaller infrastructure. What you’ll do: * Building tooling to ease the provisioning and scaling of infrastructure resources. * Continuously improve and scale infrastructure components to handle growth. * Improve overall systems performance and investigate failures taking part actively in future improvements discussion. * Ensure systems availability, reachability, and maintainability building the necessary instrumentation, tooling, and alarming systems in order to escalate abnormalities. * Being influential in monitoring and capacity planning together with the application development teams and in alignment with the business goals. What you bring in: * Extensive knowledge of system administration on Linux environments, preferably working on high throughput and low latency systems. * Strong hands-on experience with GCP services (or transferable AWS/Azure skills) — networking, IAM, compute, storage, Kubernetes (GKE/EKS/AKS). * Extensive knowledge of Docker and Kubernetes. * Excellent understanding of distributed system design across process and site boundaries. * Hands-on experience with service orchestration, management, deployment activities, configuration management and all necessary automation. * Strong grasp of process isolation and containerization concepts, being able to apply them when necessary. * Container orchestration: Deep understanding of Kubernetes — deploying, scaling, monitoring clusters. * Monitoring & Observability: Experience with tools like Prometheus, Grafana, Stackdriver, Datadog, New Relic, etc. * Incident management: Practical experience responding to incidents, performing root cause analysis, and improving system reliability. * Security best practices: Knowledge of cloud security, secrets management, and compliance basics. * Good understanding of software development lifecycle, versioning, building, testing, staging and deployment processes with a strong continuous delivery mindset. It would be great if you also have: * Experience developing kubernetes operators. * Experience deploying and scaling apache cassandra, scyllaDB, mysql, postgresql, redis or memcached. * Go programming language experience or willingness to learn coding in Go(it'll help us build new k8s operators and improve the existing ones). What we offer: We support growth through learning resources, leadership programs, mentoring, and real hands-on work. People can move between teams and projects to build new skills and keep things interesting. We offer clear internal mobility and a transparent path for progression, with leaders who stay involved and provide guidance throughout the year. In addition, you will benefit from: * A comprehensive compensation package: Learning and development allowance, voluntary provident fund (VPF) and/or national pension scheme (NPS) tax saving option provided, creche allowance * Modern tools to do your best work: Choose your preferred computer and phone within our budget, so you can work comfortably and efficiently. * A people-focused office culture: We value in-person collaboration and follow an office-first model, with some flexibility. Our offices offer a vibrant environment with opportunities to learn, connect, and recharge, from breakfast, lunch and quiet spaces to team activities such as movie nights, tech meetups, and cultural events. There's something for everyone. * Truecaller’s “Lab Days” offer a space for imagination: 5 days each quarter, where everyone steps away from their normal tasks to explore new, bold ideas and build things they’ve always wanted to. It’s a space where curiosity leads the way, and prototypes take shape. Some concepts even make it into production, and a few have grown into real features used by millions today. Lab Days allow you to be creative, learn fast, and help shape Truecaller's future. Come as you are: Truecaller is committed to building a diverse and inclusive team. We believe that a wide range of backgrounds, perspectives, and experiences strengthens our products and our culture. No matter where you're from, what language you speak, or how you identify, we value what makes you unique and would love to get to know you. Sounds like a great opportunity? We will fill the position as soon as we find the right candidate, so please send your application as soon as possible. As part of the recruitment process, we will conduct a background check. We only accept applications in English.
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. *Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role Site Reliability Engineers keep GitLab's user-facing services and production systems running reliably at scale. They combine software engineering with operational excellence, applying sound engineering principles, automation, and continuous improvement to build, operate, and evolve our production infrastructure. This is a single application for Site Reliability Engineering opportunities across our Infrastructure Platforms department. Rather than asking you to choose the right team or level upfront, we evaluate your skills holistically and match you to the opportunity that best aligns with your experience and our hiring needs. We hire Site Reliability Engineers from Intermediate through Senior Staff across multiple Infrastructure Platforms teams. We don't expect every candidate to have experience with every technology in our environment. We're looking for engineers with strong technical fundamentals, a growth mindset, and the ability to learn quickly. We'll support you in becoming successful with GitLab's tools, systems, and ways of working. How our SRE hiring works Because this is a single application for SRE roles across Infrastructure Platforms, our process is built to evaluate you once and match you well, rather than interviewing separately for every team. * Recruiter Screen: A conversation about your background, what you're looking for, and the level and teams that fit, so we can point your process in the right direction. * Core Technical: The shared assessment every SRE candidate takes, regardless of eventual team. A low-stress, collaborative discussion covering source code, system architecture, and incident review. * Peer Technical: Team-specific depth, run by SREs from the team you're most likely to join, focused on the problems that team actually works on. * Hiring Manager Interview: A conversation about ownership, judgment, execution, collaboration, and growth, the non-technical signals that make an SRE effective at GitLab. * Skip-Level Interview: A conversation with a senior leader on values alignment, and how you'll work across teams. After your interviews, we consider your performance alongside our current hiring needs to confirm the level and team where you'll do your best work. Interview results are a major factor, and final placement also reflects our active hiring priorities at the time. What level am I? We calibrate your level during the process, but here is roughly what each looks like so you know where you might land. Intermediate * You make meaningful contributions to reliability, automation, and operational efficiency, working independently within a scoped area * You diagnose issues on your own, understand system dependencies, and can explain the tradeoffs you made * You prioritize well, break work into manageable steps, and use automation to reduce toil * You document your work clearly and keep yourself moving without needing check-ins Senior * You drive reliability improvements across multiple projects or services and prioritize them based on real system needs * You lead investigations, anticipate cascading failures, and coordinate incident response * You own delivery end to end, unblock others, and improve the patterns your team works by * You communicate complex ideas clearly, influence how work gets done, and enable coordination across teams Staff * You shape reliability strategy across teams and services and define patterns that others reuse * You introduce prevention strategies, identify systemic weaknesses, and influence incident response practices beyond your immediate area * You design execution and automation approaches that work at organizational scale * You connect reliability work to platform and business needs Senior Staff * You set technical direction for reliability across a sub-department, not just a team * You drive the hardest, most ambiguous systems problems and establish standards and guardrails that multiple teams adopt * You mentor Staff and Senior engineers * You align reliability strategy with long-range platform direction and represent Infrastructure's interests across the wider Engineering organization What you'll do * Keep user-facing services and production systems reliable, scalable, and efficient * Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflows * Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling * Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps * Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately * Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outages * Take part in incident response and post-incident reviews, turning learnings into changes in automation and process * Document runbooks, architecture decisions, and reviews so your findings become repeatable practices What you'll bring * Experience keeping production systems reliable, combining an operations mindset with real software engineering practice * Experience building net-new infrastructure tooling and automation, not just configuring existing tools. For example, Terraform modules, Kubernetes operators or controllers, or production automation and services written from scratch * The ability to read, debug, and reason about code. Most of our teams work in Go; some work in Ruby. You can discuss a piece of code's behavior, performance, and failure modes * Experience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your level * Hands-on experience with at least one major cloud provider (GCP or AWS) * Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs, and using data to inform operational decisions * Comfort participating in on-call and incident response, with a structured approach to troubleshooting under pressure * Strong written communication and the ability to operate as a manager-of-one in an async, distributed environment * A track record of using automation, and increasingly AI, to reduce toil and improve how you and your team work * Alignment with GitLab's values and a commitment to working in accordance with them About the team Infrastructure Platforms is responsible for the availability, reliability, performance, and scalability of GitLab's user-facing services, most notably GitLab.com. The department spans sub-departments including Production Engineering and Dedicated, and the teams within them own everything from the production fleet and networking platform to observability, incident response, and our single-tenant Dedicated offering. We are a globally distributed, all-remote group that works asynchronously, favors automation over toil, and closes the loop with monitoring and metrics to drive accountability. For more on how we work, see the Infrastructure Handbook Page. The base salary range for this role’s listed level is currently for residents of the United States only. This range is intended to reflect the role's base salary rate in locations throughout the US. Grade level and salary ranges are determined through interviews and a review of education, experience, knowledge, skills, abilities of the applicant, equity with other team members, alignment with market data, and geographic location. The base salary range does not include any bonuses, equity, or benefits. See more information on our benefits and equity. Sales roles are also eligible for incentive pay targeted at up to 100% of the offered base salary. United States Salary Range $126,400—$314,400 USD HOW GITLAB SUPPORTS FULL-TIME EMPLOYEES * Benefits to support your health, finances, and well-being * Flexible Paid Time Off * Team Member Resource Groups * Equity Compensation & Employee Stock Purchase Plan * Growth and Development Fund * Parental Leave Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualification. If you're excited about this role, please apply and allow our recruiters to assess your application. ---------------------------------------------------------------------------------------------------------------------------------- Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process. Privacy Policy: Please review our Recruitment Privacy Policy. Your privacy is important to us. GitLab is proud to be an equal opportunity workplace and is an affirmative action employer. GitLab’s policies and practices relating to recruitment, employment, career development and advancement, promotion, and retirement are based solely on merit, regardless of race, color, religion, ancestry, sex (including pregnancy, lactation, sexual orientation, gender identity, or gender expression), national origin, age, citizenship, marital status, mental or physical disability, genetic information (including family medical history), discharge status from the military, protected veteran status (which includes disabled veterans, recently separated veterans, active duty wartime or campaign badge veterans, and Armed Forces service medal veterans), or any other basis protected by law. GitLab will not tolerate discrimination or harassment based on any of these characteristics. See also GitLab’s EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know during the recruiting process.