Amazon Development Center U.S., Inc.
Seattle, Washington
Are you excited about the opportunity to join a diverse team working together to influence the direction of an innovative new cloud database? Does contributing to product engineering and business strategy for a new distributed database service match your career goals? If so, the role of Senior Product Manager in the Aurora Distributed SQL organization may be your dream job. Key job responsibilities Your primary responsibility as a Senior Product Manager - Technical (PMT) is to help accelerate the product engineering and business growth of Amazon Aurora DSQL. A Senior PMT achieves this growth by shaping and leading product strategy - turning insights and customer feedback into plans that accelerate delivery and delight customers. This includes taking ownership of roadmap prioritization, and contributing to feature development by acting as the voice of the customer throughout the software development lifecycle. You will collaborate directly with software engineering teams, and interact with senior business partners across marketing, finance, and sales organizations. A day in the life A successful candidate will demonstrate proven product management skills with experience delivering features for large scale enterprise information systems. The candidate will be comfortable managing competing priorities and be able to bring clarity and order to ambiguous scenarios. The candidate will demonstrate subject matter expertise in database technology and is familiar with related storage and server hardware. Exemplary written and verbal communication skills are important, as the candidate will frequently work with leadership and colleagues that cross organizational boundaries. BASIC QUALIFICATIONS - 5+ years of technical product management with internet business experience - 3+ years of technical (software development, network development, IT, other related) experience - Experience in taking a product from conception & definition phase through engineering design and taking it to market - Experience delivering large-scale SaaS, PaaS or LaaS products where you are responsible for the full product lifecycle, from concept through GTM (go to market) - 3+ years of database (eg. SQL, NoSQL, Hadoop, Spark, Kafka, Kinesis) experience, or degree in advanced technology PREFERRED QUALIFICATIONS - Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations - Experience working within teams delivering software products and features using agile methodologies Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, WA, Seattle - 152 900.00 USD annually
09/23/2026
Full time
Are you excited about the opportunity to join a diverse team working together to influence the direction of an innovative new cloud database? Does contributing to product engineering and business strategy for a new distributed database service match your career goals? If so, the role of Senior Product Manager in the Aurora Distributed SQL organization may be your dream job. Key job responsibilities Your primary responsibility as a Senior Product Manager - Technical (PMT) is to help accelerate the product engineering and business growth of Amazon Aurora DSQL. A Senior PMT achieves this growth by shaping and leading product strategy - turning insights and customer feedback into plans that accelerate delivery and delight customers. This includes taking ownership of roadmap prioritization, and contributing to feature development by acting as the voice of the customer throughout the software development lifecycle. You will collaborate directly with software engineering teams, and interact with senior business partners across marketing, finance, and sales organizations. A day in the life A successful candidate will demonstrate proven product management skills with experience delivering features for large scale enterprise information systems. The candidate will be comfortable managing competing priorities and be able to bring clarity and order to ambiguous scenarios. The candidate will demonstrate subject matter expertise in database technology and is familiar with related storage and server hardware. Exemplary written and verbal communication skills are important, as the candidate will frequently work with leadership and colleagues that cross organizational boundaries. BASIC QUALIFICATIONS - 5+ years of technical product management with internet business experience - 3+ years of technical (software development, network development, IT, other related) experience - Experience in taking a product from conception & definition phase through engineering design and taking it to market - Experience delivering large-scale SaaS, PaaS or LaaS products where you are responsible for the full product lifecycle, from concept through GTM (go to market) - 3+ years of database (eg. SQL, NoSQL, Hadoop, Spark, Kafka, Kinesis) experience, or degree in advanced technology PREFERRED QUALIFICATIONS - Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations - Experience working within teams delivering software products and features using agile methodologies Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, WA, Seattle - 152 900.00 USD annually
Gravitee is a 2025 Gartner Magic Quadrant Leader , on a mission to govern the world's intelligence . We deliver the industry's most advanced platform for Any API, Any Event, and Any AI Agent , trusted by global leaders like Michelin, Roche, and Blue Yonder. Why join us? The Mission : We are the first to bridge traditional API Management with the new frontier of AI Agent Security The Momentum : A high-growth Leader - combining market credibility with startup speed The DNA : We hire people who Hold Nothing Back - passionate builders who want to redefine digital infrastructure Don't just watch the AI revolution. Build the infrastructure that controls and secures it. The Role We are looking for a Senior Software Engineer to build and maintain the identity and authorization features of Gravitee Access Management (AM) - across the AM runtime and the access-management experience in Gamma, Gravitee's next generation product surface. This is a new role. Today, AM engineering is based entirely in Europe. This hire establishes US-hours ownership of Level 3 and Level 4 authentication and authorization incidents, and adds delivery capacity toward AM parity in Gamma - part of building sustainable L3/L4 engineering capability in the US. You will split your time roughly 80% feature delivery and 20% L3/L4 support and bug fixing (it varies week to week), working as an embedded member of the AM team, which is based in Europe. What You Will Be Doing In this role, you will: Design and deliver features end to end, from discovery and technical design through implementation, testing, release, and iteration. Build and maintain identity and authorization features of Gravitee Access Management, across the AM runtime and the AM experience in Gamma. Implement and support OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0 as both IdP and SP, SCIM, and FAPI/CIBA/UMA profiles. Work with token and session semantics - JWT, JWKS, key rotation, revocation, introspection, MFA and step-up, WebAuthn/FIDO2, and IdP federation and social login. Keep security behavior and upgrades safe: standards compliance, secure defaults, certificate and secret handling, consent, audit logs, and defenses against token replay, SSRF, and account takeover. Own safe migrations and backward compatibility across MongoDB and JDBC, and support multi-domain, multi-region deployments and login/token endpoint performance. Own US-hours Level 3 and Level 4 escalations for AM customers as part of the L3 pager duty rotation. Use LLMs and AI-assisted development tools thoughtfully for prototyping, implementation, testing, debugging, and exploration, applying sound engineering judgment to validate AI-generated work. Write meaningful automated tests and contribute to reliable delivery practices. • Collaborate with product managers, designers, engineers, and technical leaders - including the AM team based in Europe - to discover effective solutions and improve them through code and design reviews. Share what you learn and help the team make practical choices as identity standards and protocols evolve. Essential Skills We are looking for evidence that you can succeed in the role, whether gained through employment, open-source work, or equivalent practical experience: 5+ years building and running production backend software, on a team that ships and supports its own product; you have personally resolved production incidents. Strong Java experience (C# accepted if the object-oriented depth is there), with Maven and a reactive stack such as Vert.x/RxJava. Deep working knowledge of identity standards: OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0, SCIM, and FAPI/CIBA/UMA profiles. • Solid grasp of token and session semantics: JWT, JWKS, rotation, revocation, introspection, MFA/step-up, WebAuthn/FIDO2, and IdP federation. A security-first mindset: secure defaults, certificate and secret handling, audit logging, and awareness of token replay, SSRF, and account-takeover risks. Experience with safe migrations and backward compatibility across persistent data stores such as MongoDB or JDBCbacked relational databases. Git-based workflow, code review, and writing your own automated tests. Hands-on experience using LLMs or AI coding assistants as part of an engineering workflow, combined with the judgment to review and improve their output. Clear communication, collaborative problem-solving, and the ability to take an ambiguous problem through to production. Desired Skills You do not need to match every item. We would be especially interested in experience with: • Experience at an API gateway, proxy, or service-mesh vendor, or on the API platform team of a large company (e.g., Kong, Google Apigee, MuleSoft, Tyk, Solo.io, Traefik, WSO2). Kubernetes operators and CRDs; OpenAPI tooling; service mesh or Envoy experience. Docker, Kubernetes, and cloud-native application delivery. Model Context Protocol (MCP), Agent2Agent (A2A), tool calling, LLM proxies, or other emerging AI protocols and standards. Prior production experience is not required. Building or operating LLM-powered applications, RAG systems, or agentic workflows - especially their security, governance, and observability needs. Open-source software or enterprise developer platforms. Who Thrives at Gravitee Our growth is powered by people who bring passion to what they build, professionalism to how they work, and a commitment to doing things well. You will thrive here if you: • Bring energy and a constructive attitude to the team. • Adapt quickly and enjoy learning unfamiliar technologies and domains. • Take ownership, communicate clearly, and follow through with urgency. • Balance delivery speed with thoughtful engineering judgment. • Start with the customer problem and care about the quality of the experience you create. • Enjoy working in a fast-moving, collaborative, international environment. Life at Gravitee At Gravitee, we invest in humans, not just roles. You'll get: • Salary of $160,000 • Competitive medical coverage. • Pension / 401(k) program options. • Stock options - you build it, you own it. • 25 days of holiday plus in-country national holidays. • Three mental health days and a wellness allowance. • Your birthday off. • A professional development budget to support your growth. • A hybrid work culture with hubs across regions. • Quarterly team events and an annual company offsite. • A collaborative, international company culture. • Opportunities to grow your scope and career as Gravitee grows. At Gravitee, we believe diverse perspectives make better products and stronger teams. No employee or applicant will be treated less favorably on the grounds of sex, marital status, race, color, nationality, ethnic or national origin, disability, gender, sexual orientation, gender identity, age, pregnancy or maternity, marital or civil partner status, religion, or belief. By applying, you consent to Gravitee storing and processing the personal information you submit as part of the recruitment process.
09/23/2026
Full time
Gravitee is a 2025 Gartner Magic Quadrant Leader , on a mission to govern the world's intelligence . We deliver the industry's most advanced platform for Any API, Any Event, and Any AI Agent , trusted by global leaders like Michelin, Roche, and Blue Yonder. Why join us? The Mission : We are the first to bridge traditional API Management with the new frontier of AI Agent Security The Momentum : A high-growth Leader - combining market credibility with startup speed The DNA : We hire people who Hold Nothing Back - passionate builders who want to redefine digital infrastructure Don't just watch the AI revolution. Build the infrastructure that controls and secures it. The Role We are looking for a Senior Software Engineer to build and maintain the identity and authorization features of Gravitee Access Management (AM) - across the AM runtime and the access-management experience in Gamma, Gravitee's next generation product surface. This is a new role. Today, AM engineering is based entirely in Europe. This hire establishes US-hours ownership of Level 3 and Level 4 authentication and authorization incidents, and adds delivery capacity toward AM parity in Gamma - part of building sustainable L3/L4 engineering capability in the US. You will split your time roughly 80% feature delivery and 20% L3/L4 support and bug fixing (it varies week to week), working as an embedded member of the AM team, which is based in Europe. What You Will Be Doing In this role, you will: Design and deliver features end to end, from discovery and technical design through implementation, testing, release, and iteration. Build and maintain identity and authorization features of Gravitee Access Management, across the AM runtime and the AM experience in Gamma. Implement and support OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0 as both IdP and SP, SCIM, and FAPI/CIBA/UMA profiles. Work with token and session semantics - JWT, JWKS, key rotation, revocation, introspection, MFA and step-up, WebAuthn/FIDO2, and IdP federation and social login. Keep security behavior and upgrades safe: standards compliance, secure defaults, certificate and secret handling, consent, audit logs, and defenses against token replay, SSRF, and account takeover. Own safe migrations and backward compatibility across MongoDB and JDBC, and support multi-domain, multi-region deployments and login/token endpoint performance. Own US-hours Level 3 and Level 4 escalations for AM customers as part of the L3 pager duty rotation. Use LLMs and AI-assisted development tools thoughtfully for prototyping, implementation, testing, debugging, and exploration, applying sound engineering judgment to validate AI-generated work. Write meaningful automated tests and contribute to reliable delivery practices. • Collaborate with product managers, designers, engineers, and technical leaders - including the AM team based in Europe - to discover effective solutions and improve them through code and design reviews. Share what you learn and help the team make practical choices as identity standards and protocols evolve. Essential Skills We are looking for evidence that you can succeed in the role, whether gained through employment, open-source work, or equivalent practical experience: 5+ years building and running production backend software, on a team that ships and supports its own product; you have personally resolved production incidents. Strong Java experience (C# accepted if the object-oriented depth is there), with Maven and a reactive stack such as Vert.x/RxJava. Deep working knowledge of identity standards: OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0, SCIM, and FAPI/CIBA/UMA profiles. • Solid grasp of token and session semantics: JWT, JWKS, rotation, revocation, introspection, MFA/step-up, WebAuthn/FIDO2, and IdP federation. A security-first mindset: secure defaults, certificate and secret handling, audit logging, and awareness of token replay, SSRF, and account-takeover risks. Experience with safe migrations and backward compatibility across persistent data stores such as MongoDB or JDBCbacked relational databases. Git-based workflow, code review, and writing your own automated tests. Hands-on experience using LLMs or AI coding assistants as part of an engineering workflow, combined with the judgment to review and improve their output. Clear communication, collaborative problem-solving, and the ability to take an ambiguous problem through to production. Desired Skills You do not need to match every item. We would be especially interested in experience with: • Experience at an API gateway, proxy, or service-mesh vendor, or on the API platform team of a large company (e.g., Kong, Google Apigee, MuleSoft, Tyk, Solo.io, Traefik, WSO2). Kubernetes operators and CRDs; OpenAPI tooling; service mesh or Envoy experience. Docker, Kubernetes, and cloud-native application delivery. Model Context Protocol (MCP), Agent2Agent (A2A), tool calling, LLM proxies, or other emerging AI protocols and standards. Prior production experience is not required. Building or operating LLM-powered applications, RAG systems, or agentic workflows - especially their security, governance, and observability needs. Open-source software or enterprise developer platforms. Who Thrives at Gravitee Our growth is powered by people who bring passion to what they build, professionalism to how they work, and a commitment to doing things well. You will thrive here if you: • Bring energy and a constructive attitude to the team. • Adapt quickly and enjoy learning unfamiliar technologies and domains. • Take ownership, communicate clearly, and follow through with urgency. • Balance delivery speed with thoughtful engineering judgment. • Start with the customer problem and care about the quality of the experience you create. • Enjoy working in a fast-moving, collaborative, international environment. Life at Gravitee At Gravitee, we invest in humans, not just roles. You'll get: • Salary of $160,000 • Competitive medical coverage. • Pension / 401(k) program options. • Stock options - you build it, you own it. • 25 days of holiday plus in-country national holidays. • Three mental health days and a wellness allowance. • Your birthday off. • A professional development budget to support your growth. • A hybrid work culture with hubs across regions. • Quarterly team events and an annual company offsite. • A collaborative, international company culture. • Opportunities to grow your scope and career as Gravitee grows. At Gravitee, we believe diverse perspectives make better products and stronger teams. No employee or applicant will be treated less favorably on the grounds of sex, marital status, race, color, nationality, ethnic or national origin, disability, gender, sexual orientation, gender identity, age, pregnancy or maternity, marital or civil partner status, religion, or belief. By applying, you consent to Gravitee storing and processing the personal information you submit as part of the recruitment process.
NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Senior Oracle EPM Financial Consolidation and Close (FCCS) Production Support Specialist (REMOTE) to join our team! This is a fully remote role and can be based anywhere in the US. We are seeking an experienced Oracle EDMCS professional to lead the implementation, maintenance, and governance of enterprise metadata and master data management solutions. You will ensure that hierarchies, dimensions, and reference data remain accurate, synchronized, and well-controlled across Oracle EPM, ERP, and downstream third-party applications. This role partners with Finance, Accounting, and IT stakeholders to support, enhance, and optimize existing financial systems and integrations. Job Responsibilities Include: Provide day-to-day functional and technical support for Oracle Financial Consolidation and Close (FCCS), including metadata management, consolidation rules, workflows, and process monitoring. Configure core EDMCS components including Node Types, Node Sets, Viewpoints, Hierarchies, and Property Definitions Design scalable organizational data structures for the Chart of Accounts (COA), Cost Centers, Departments, Legal Entities, and Product Hierarchies Set up integrations and data synchronization between EDMCS and Oracle Cloud applications (such as FCCS, Planning/Freeform, and ERP Cloud) Maintain and enhance ERP, EPM , Custom applications scenarios, entities, accounts, and intercompany structures. Define and manage metadata approval workflows, request interfaces, and change control mechanisms. Establish business validation rules, derivations, and subscriptions to automate and control data changes. Maintain system security, access controls, and granular role assignments Troubleshoot consolidation issues, data discrepancies, and financial reporting variances across source systems. Manage system operations related to the close process, including running consolidations, translations, calculations, and validations. Collaborate with IT, data engineering, and application owners during integration projects, upgrades, and system migrations. Serve as a subject matter expert (SME) on EDMCS functionality and work with stakeholder for any BAU changes Ensure reporting aligns with corporate accounting policies, financial structures, and audit requirements. Provide user training, guidance, and support for change request submission and hierarchy management. Document issues, enhancements, and process changes following ITIL standards (ServiceNow or similar). Support internal and external audit activities through documentation, controls, and system evidence. Recommend technology and process improvements that enhance financial operations efficiency. Work with cross functional teams to automate manual financial workflows using the appropriate tools. Identify opportunities to optimize metadata structures, automation, and data governance workflows. Basic Qualifications: 5-8+ years of experience in Oracle EPM/ERP, with hands-on expertise specifically in Oracle EDMCS. 5+ years of experience supporting Oracle FCCS and/or Hyperion Financial Management. 3-5 years of experience with data exchange/FDMEE/FDM data integrations. 3-5 years of strong understanding of financial accounting, consolidations, intercompany eliminations, FX translation, and period-close processes. + 5 years of experience with Oracle Cloud Narrative Reporting/Reports and/or Financial Reporting Studio Web +5 years of experience with Oracle Smart View (creating new reports, supporting ad hoc, break fix, etc.) 5+ years of Experience managing hierarchies, dimensions, and metadata governance processes. 5+ years of Hands-on experience configuring integrations, exports/imports, and connections to source/target systems. Preferred Skills: Strong knowledge of automation using Oracle EPMAutomate. Strong analytical, documentation, and problem solving skills. Highly organized and detail oriented. Strong communicator able to partner effectively with Finance, IT, and business stakeholders. Adaptable and proactive, with a continuous improvement mindset. Comfortable managing multiple priorities in a dynamic environment. Experience with Oracle ERP Cloud, data warehouses, OCI, Oracle API Gateway. Basic understanding on Oracle ARCS, ,FreeForm, TRCS products knowledge. Experience in the Banking sector is preferred. Ability to write Groovy rules, REST API calls, or SQL is preferred. Oracle EPM Cloud EDMCS certification is preferred. NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this role is $123,210 - $205,350 per annum. Actual compensation will depend on a number of factors, including the candidate's relevant experience, technical skills, and other qualifications. This position may also be eligible for incentive compensation based on individual and/or company performance. This position is eligible for company benefits including medical, dental, and vision insurance with an employer contribution, flexible spending or health savings account, life and AD&D insurance, short and long term disability coverage, paid time off, employee assistance, participation in a 401k program with company match, and additional voluntary or legally-required benefits. About NTT DATA NTT DATA is a $30 billion business and technology services leader, serving 75% of the Fortune Global 100. We are committed to accelerating client success and positively impacting society through responsible innovation. We are one of the world's leading AI and digital infrastructure providers, with unmatched capabilities in enterprise-scale AI, cloud, security, connectivity, data centers and application services. our consulting and Industry solutions help organizations and society move confidently and sustainably into the digital future. As a Global Top Employer, we have experts in more than 50 countries. We also offer clients access to a robust ecosystem of innovation centers as well as established and start-up partners. NTT DATA is a part of NTT Group, which invests over $3 billion each year in R&D. Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored to each client's needs. While many positions offer remote or hybrid work options, these arrangements are subject to change based on client requirements. For employees near an NTT DATA office or client site, in-office attendance may be required for meetings or events, depending on business needs. At NTT DATA, we are committed to staying flexible and meeting the evolving needs of both our clients and employees. NTT DATA recruiters will never ask for payment or banking information and will only email addresses. If you are requested to provide payment or disclose banking information, please submit a contact us form, NTT DATA endeavors to make accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact us at This contact information is for accommodation requests only and cannot be used to inquire about the status of applications. NTT DATA is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status. For our EEO Policy Statement, please click here. If you'd like more information on your EEO rights under the law, please click here. For Pay Transparency information, please click here.
09/23/2026
Full time
NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Senior Oracle EPM Financial Consolidation and Close (FCCS) Production Support Specialist (REMOTE) to join our team! This is a fully remote role and can be based anywhere in the US. We are seeking an experienced Oracle EDMCS professional to lead the implementation, maintenance, and governance of enterprise metadata and master data management solutions. You will ensure that hierarchies, dimensions, and reference data remain accurate, synchronized, and well-controlled across Oracle EPM, ERP, and downstream third-party applications. This role partners with Finance, Accounting, and IT stakeholders to support, enhance, and optimize existing financial systems and integrations. Job Responsibilities Include: Provide day-to-day functional and technical support for Oracle Financial Consolidation and Close (FCCS), including metadata management, consolidation rules, workflows, and process monitoring. Configure core EDMCS components including Node Types, Node Sets, Viewpoints, Hierarchies, and Property Definitions Design scalable organizational data structures for the Chart of Accounts (COA), Cost Centers, Departments, Legal Entities, and Product Hierarchies Set up integrations and data synchronization between EDMCS and Oracle Cloud applications (such as FCCS, Planning/Freeform, and ERP Cloud) Maintain and enhance ERP, EPM , Custom applications scenarios, entities, accounts, and intercompany structures. Define and manage metadata approval workflows, request interfaces, and change control mechanisms. Establish business validation rules, derivations, and subscriptions to automate and control data changes. Maintain system security, access controls, and granular role assignments Troubleshoot consolidation issues, data discrepancies, and financial reporting variances across source systems. Manage system operations related to the close process, including running consolidations, translations, calculations, and validations. Collaborate with IT, data engineering, and application owners during integration projects, upgrades, and system migrations. Serve as a subject matter expert (SME) on EDMCS functionality and work with stakeholder for any BAU changes Ensure reporting aligns with corporate accounting policies, financial structures, and audit requirements. Provide user training, guidance, and support for change request submission and hierarchy management. Document issues, enhancements, and process changes following ITIL standards (ServiceNow or similar). Support internal and external audit activities through documentation, controls, and system evidence. Recommend technology and process improvements that enhance financial operations efficiency. Work with cross functional teams to automate manual financial workflows using the appropriate tools. Identify opportunities to optimize metadata structures, automation, and data governance workflows. Basic Qualifications: 5-8+ years of experience in Oracle EPM/ERP, with hands-on expertise specifically in Oracle EDMCS. 5+ years of experience supporting Oracle FCCS and/or Hyperion Financial Management. 3-5 years of experience with data exchange/FDMEE/FDM data integrations. 3-5 years of strong understanding of financial accounting, consolidations, intercompany eliminations, FX translation, and period-close processes. + 5 years of experience with Oracle Cloud Narrative Reporting/Reports and/or Financial Reporting Studio Web +5 years of experience with Oracle Smart View (creating new reports, supporting ad hoc, break fix, etc.) 5+ years of Experience managing hierarchies, dimensions, and metadata governance processes. 5+ years of Hands-on experience configuring integrations, exports/imports, and connections to source/target systems. Preferred Skills: Strong knowledge of automation using Oracle EPMAutomate. Strong analytical, documentation, and problem solving skills. Highly organized and detail oriented. Strong communicator able to partner effectively with Finance, IT, and business stakeholders. Adaptable and proactive, with a continuous improvement mindset. Comfortable managing multiple priorities in a dynamic environment. Experience with Oracle ERP Cloud, data warehouses, OCI, Oracle API Gateway. Basic understanding on Oracle ARCS, ,FreeForm, TRCS products knowledge. Experience in the Banking sector is preferred. Ability to write Groovy rules, REST API calls, or SQL is preferred. Oracle EPM Cloud EDMCS certification is preferred. NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this role is $123,210 - $205,350 per annum. Actual compensation will depend on a number of factors, including the candidate's relevant experience, technical skills, and other qualifications. This position may also be eligible for incentive compensation based on individual and/or company performance. This position is eligible for company benefits including medical, dental, and vision insurance with an employer contribution, flexible spending or health savings account, life and AD&D insurance, short and long term disability coverage, paid time off, employee assistance, participation in a 401k program with company match, and additional voluntary or legally-required benefits. About NTT DATA NTT DATA is a $30 billion business and technology services leader, serving 75% of the Fortune Global 100. We are committed to accelerating client success and positively impacting society through responsible innovation. We are one of the world's leading AI and digital infrastructure providers, with unmatched capabilities in enterprise-scale AI, cloud, security, connectivity, data centers and application services. our consulting and Industry solutions help organizations and society move confidently and sustainably into the digital future. As a Global Top Employer, we have experts in more than 50 countries. We also offer clients access to a robust ecosystem of innovation centers as well as established and start-up partners. NTT DATA is a part of NTT Group, which invests over $3 billion each year in R&D. Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored to each client's needs. While many positions offer remote or hybrid work options, these arrangements are subject to change based on client requirements. For employees near an NTT DATA office or client site, in-office attendance may be required for meetings or events, depending on business needs. At NTT DATA, we are committed to staying flexible and meeting the evolving needs of both our clients and employees. NTT DATA recruiters will never ask for payment or banking information and will only email addresses. If you are requested to provide payment or disclose banking information, please submit a contact us form, NTT DATA endeavors to make accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact us at This contact information is for accommodation requests only and cannot be used to inquire about the status of applications. NTT DATA is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status. For our EEO Policy Statement, please click here. If you'd like more information on your EEO rights under the law, please click here. For Pay Transparency information, please click here.
We are looking for a Principal Modern Planner to own demand planning and forecasting for networking hardware supporting large-scale cloud data center builds and expansions. This role will connect data center build plans, network architecture, product roadmaps, engineering changes, BOMs, and historical demand to a long-range view of hardware requirements. The planner will work across Networking Engineering, Network Architecture, Data Center Planning, Product, Supply Chain, Procurement, Finance, and Data/Analytics to understand the underlying demand drivers, reconcile different inputs, and improve the quality and consistency of the demand plan. The role also has a strong analytics and data component. The successful candidate will use forecasting, statistical analysis, data modeling, automation, and AI/ML to improve planning processes, identify issues in the demand signal, and provide better visibility into future capacity and supply requirements. This is a hands-on Principal role for someone who can move between planning strategy, detailed data analysis, and networking hardware fundamentals, and who can work effectively with both technical teams and senior leadership. Responsibilities Own demand planning and forecasting for networking hardware supporting data center builds, expansions, and ongoing capacity requirements. Develop and maintain long-range demand forecasts, including 24+ month outlooks, across networking products, platforms, and SKUs. Work with Data Center Planning, Network Engineering, Product, Supply Chain, Procurement, and Finance to understand upcoming builds, deployment timing, architecture changes, and other demand drivers. Translate data center build plans and network architecture requirements into hardware and component demand, including the relationship between capacity, racks, network topology, BOMs, and SKUs. Develop forecasting models using historical demand, deployment trends, engineering inputs, product roadmaps, and other relevant signals. Build scenarios to understand the demand impact of changes in data center build timing, network architecture, product transitions, capacity plans, and supply constraints. Work with engineering teams to incorporate BOM changes, new product introductions, product transitions, substitutions, and EOL/EOS plans into the forecast. Identify issues in the demand signal, including double counting, overlapping assumptions, missing requirements, and inconsistent inputs across planning processes. Establish metrics and analytical methods to measure forecast accuracy, bias, volatility, and confidence, and use those insights to improve the planning process. Build and maintain the data and analytical foundation needed to connect data center plans, deployment information, engineering/BOM data, demand, supply, and inventory. Automate recurring planning, reconciliation, and reporting processes using SQL, Python, and other analytical technologies. Apply statistical forecasting, machine learning, optimization, simulation, and AI where they can materially improve planning quality or reduce manual work. Establish a regular planning cadence with engineering and business partners, including mechanisms for reviewing assumptions, reconciling changes, and obtaining alignment on the demand plan. Prepare analysis and recommendations for senior leadership on demand changes, capacity requirements, supply risks, and key planning assumptions. Lead complex planning issues across organizational boundaries and drive them to resolution. Mentor other planners and analytical team members and help establish scalable planning practices. Required Qualifications Bachelor's or Master's degree in Engineering, Computer Science, Data Science, Statistics, Operations Research, Supply Chain, Economics, or a related field. 5-8 years of experience in demand planning, forecasting, capacity planning, supply-chain planning, analytics, operations research, or a related area. Experience working with technology hardware, networking, semiconductor, cloud infrastructure, or data center infrastructure. Strong understanding of demand forecasting and long-range planning, including forecast accuracy, bias, scenario planning, and demand drivers. Experience working with complex hardware products, including BOMs, SKUs, product lifecycle, NPI, EOL/EOS, and product transitions. Experience connecting engineering, deployment, or infrastructure plans to hardware demand. Strong analytical and quantitative skills, including hands-on experience with SQL and Python/R. Experience working with large datasets and using data to investigate problems and support planning decisions. Strong cross-functional communication skills and experience working with engineering, supply chain, product, and business stakeholders. Ability to operate independently, navigate ambiguity, and influence decisions across organizations. Preferred Qualifications Experience with data center build and deployment planning or cloud infrastructure capacity planning. Experience with networking hardware such as Ethernet switches, NICs, DPUs/SmartNICs, optical transceivers, cables, or related components. Familiarity with data center networking architectures and high-performance/AI networking. Experience developing models that connect data center builds and network architecture to BOM and SKU-level demand. Experience with networking or semiconductor supply chains, including lead times, constraints, allocation, substitutions, and technology transitions. Experience with time-series forecasting, probabilistic forecasting, machine learning, optimization, simulation, or other advanced analytical methods. Experience building data pipelines, analytical datasets, dashboards, or planning tools. Experience using GenAI or automation to improve planning and forecasting processes. Experience presenting planning analysis and recommendations to senior engineering or business leadership. Qualifications Disclaimer: Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements. Range and benefit information provided in this posting are specific to the stated locations only US: Hiring Range in USD from: $90,100 to $209,500 per annum. May be eligible for bonus and equity. Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following: 1. Medical, dental, and vision insurance, including expert medical opinion 2. Short term disability and long term disability 3. Life insurance and AD&D 4. Supplemental life insurance (Employee/Spouse/Child) 5. Health care and dependent care Flexible Spending Accounts 6. Pre-tax commuter and parking benefits 7. 401(k) Savings and Investment Plan with company match 8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation. 9. 11 paid holidays 10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours. 11. Paid parental leave 12. Adoption assistance 13. Employee Stock Purchase Plan 14. Financial planning and group legal 15. Voluntary benefits including auto, homeowner and pet insurance The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted. As part of Oracle's onboarding process and consistent with applicable law, US-based employees are required to complete identity verification, which involves the collection and processing of their biometric information. Accommodations to this requirement may be granted following an individualized assessment. Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We're committed to including people with disabilities at all stages of the employment process . click apply for full job details
09/23/2026
Full time
We are looking for a Principal Modern Planner to own demand planning and forecasting for networking hardware supporting large-scale cloud data center builds and expansions. This role will connect data center build plans, network architecture, product roadmaps, engineering changes, BOMs, and historical demand to a long-range view of hardware requirements. The planner will work across Networking Engineering, Network Architecture, Data Center Planning, Product, Supply Chain, Procurement, Finance, and Data/Analytics to understand the underlying demand drivers, reconcile different inputs, and improve the quality and consistency of the demand plan. The role also has a strong analytics and data component. The successful candidate will use forecasting, statistical analysis, data modeling, automation, and AI/ML to improve planning processes, identify issues in the demand signal, and provide better visibility into future capacity and supply requirements. This is a hands-on Principal role for someone who can move between planning strategy, detailed data analysis, and networking hardware fundamentals, and who can work effectively with both technical teams and senior leadership. Responsibilities Own demand planning and forecasting for networking hardware supporting data center builds, expansions, and ongoing capacity requirements. Develop and maintain long-range demand forecasts, including 24+ month outlooks, across networking products, platforms, and SKUs. Work with Data Center Planning, Network Engineering, Product, Supply Chain, Procurement, and Finance to understand upcoming builds, deployment timing, architecture changes, and other demand drivers. Translate data center build plans and network architecture requirements into hardware and component demand, including the relationship between capacity, racks, network topology, BOMs, and SKUs. Develop forecasting models using historical demand, deployment trends, engineering inputs, product roadmaps, and other relevant signals. Build scenarios to understand the demand impact of changes in data center build timing, network architecture, product transitions, capacity plans, and supply constraints. Work with engineering teams to incorporate BOM changes, new product introductions, product transitions, substitutions, and EOL/EOS plans into the forecast. Identify issues in the demand signal, including double counting, overlapping assumptions, missing requirements, and inconsistent inputs across planning processes. Establish metrics and analytical methods to measure forecast accuracy, bias, volatility, and confidence, and use those insights to improve the planning process. Build and maintain the data and analytical foundation needed to connect data center plans, deployment information, engineering/BOM data, demand, supply, and inventory. Automate recurring planning, reconciliation, and reporting processes using SQL, Python, and other analytical technologies. Apply statistical forecasting, machine learning, optimization, simulation, and AI where they can materially improve planning quality or reduce manual work. Establish a regular planning cadence with engineering and business partners, including mechanisms for reviewing assumptions, reconciling changes, and obtaining alignment on the demand plan. Prepare analysis and recommendations for senior leadership on demand changes, capacity requirements, supply risks, and key planning assumptions. Lead complex planning issues across organizational boundaries and drive them to resolution. Mentor other planners and analytical team members and help establish scalable planning practices. Required Qualifications Bachelor's or Master's degree in Engineering, Computer Science, Data Science, Statistics, Operations Research, Supply Chain, Economics, or a related field. 5-8 years of experience in demand planning, forecasting, capacity planning, supply-chain planning, analytics, operations research, or a related area. Experience working with technology hardware, networking, semiconductor, cloud infrastructure, or data center infrastructure. Strong understanding of demand forecasting and long-range planning, including forecast accuracy, bias, scenario planning, and demand drivers. Experience working with complex hardware products, including BOMs, SKUs, product lifecycle, NPI, EOL/EOS, and product transitions. Experience connecting engineering, deployment, or infrastructure plans to hardware demand. Strong analytical and quantitative skills, including hands-on experience with SQL and Python/R. Experience working with large datasets and using data to investigate problems and support planning decisions. Strong cross-functional communication skills and experience working with engineering, supply chain, product, and business stakeholders. Ability to operate independently, navigate ambiguity, and influence decisions across organizations. Preferred Qualifications Experience with data center build and deployment planning or cloud infrastructure capacity planning. Experience with networking hardware such as Ethernet switches, NICs, DPUs/SmartNICs, optical transceivers, cables, or related components. Familiarity with data center networking architectures and high-performance/AI networking. Experience developing models that connect data center builds and network architecture to BOM and SKU-level demand. Experience with networking or semiconductor supply chains, including lead times, constraints, allocation, substitutions, and technology transitions. Experience with time-series forecasting, probabilistic forecasting, machine learning, optimization, simulation, or other advanced analytical methods. Experience building data pipelines, analytical datasets, dashboards, or planning tools. Experience using GenAI or automation to improve planning and forecasting processes. Experience presenting planning analysis and recommendations to senior engineering or business leadership. Qualifications Disclaimer: Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements. Range and benefit information provided in this posting are specific to the stated locations only US: Hiring Range in USD from: $90,100 to $209,500 per annum. May be eligible for bonus and equity. Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following: 1. Medical, dental, and vision insurance, including expert medical opinion 2. Short term disability and long term disability 3. Life insurance and AD&D 4. Supplemental life insurance (Employee/Spouse/Child) 5. Health care and dependent care Flexible Spending Accounts 6. Pre-tax commuter and parking benefits 7. 401(k) Savings and Investment Plan with company match 8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation. 9. 11 paid holidays 10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours. 11. Paid parental leave 12. Adoption assistance 13. Employee Stock Purchase Plan 14. Financial planning and group legal 15. Voluntary benefits including auto, homeowner and pet insurance The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted. As part of Oracle's onboarding process and consistent with applicable law, US-based employees are required to complete identity verification, which involves the collection and processing of their biometric information. Accommodations to this requirement may be granted following an individualized assessment. Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We're committed to including people with disabilities at all stages of the employment process . click apply for full job details
ASRC Federal is a leading government contractor furthering missions in space, public health and defense. As an Alaska Native owned corporation, our work helps secure an enduring future for our shareholders. Join our team and discover why we are a top veteran employer and Certified Great Place to Work ASRC Federal Data Networks Corporation (DNC) is seeking a senior, versatile, Systems Engineer to serve as the critical technical integrator within the Traffic Coordination System for Space (TraCSS) Scaled Agile Framework (SAFe) ecosystem. Operating at the Release and Requirements Management level, this role bridges the gap between high-level government strategy and multi-vendor execution. Our contract team supports the Office of Space Commerce, Space Operations Division, TraCSS Engineering Branch located in Suitland MD. The successful candidate will take strategic direction from the Government Business Owner, Program Management, and Product Owners to ensure systemic technical integrity across the TraCSS enterprise. Acting as a collaborative technical leader, you will work hand-in-hand with the System Integrator's Release Train Engineer (RTE) and coordinate directly with the System Integrator and Presentation Layer Scrum Masters to clear program-level technical roadblocks, refine backlogs, and maintain a robust architectural runway. Key Responsibilities: SAFe Technical Leadership: Own the technical "big picture" for system integration across the Agile Release Train (ART). Partner with the System Integrator's RTE to ensure multi-vendor scrum teams remain aligned with the broader product vision. Cross-Vendor Integration & Synchronization: Coordinate directly with the Presentation Layer and System Integrator contractors, satellite Owner/Operators (O/Os), and commercial vendors to identify, analyze, and resolve complex engineering and integration issues. Requirements & Backlog Refinement: Collaborate closely with the TraCSS Requirements Engineer to translate high-level stakeholder needs into technical clarity. Assist Product Owners in Epic/Story creation, technical dependency mapping, and backlog refinement. Interagency & Industry Liaison: Represent the Systems Engineering integrated program team in technical discussions and meetings with commercial space situational awareness providers, civil operators, and external federal agencies (including NASA and industry partners) to track studies and align data standards. Mandatory Qualifications & Criteria: 1. Agile & Framework Proficiency SAFe Ecosystems : Proven experience operating as a System Engineer, Solution Architect, or Product Owner within a Scaled Agile Framework (SAFe) environment or on a large-scale Agile Release Train (ART). Interface Coordination : Documented success bridging gaps between product management levels and technical Scrum Masters/Developers in a multi-contractor configuration. Tooling : Proficiency with the Atlassian Suite (Jira, Jira Service Management, Confluence) for tracking requirements, dependencies, and program risks. 2. Compliance & Education Education : Bachelor's degree in Aerospace Engineering, Physics, Mathematics, Computer Science, or a related technical field. Background Processing : Must successfully pass NOAA background screening requirements and sign a mandatory Office of Space Commerce Non-Disclosure Agreement (NDA). Citizenship : US Citizenship required. Highly Desired 1. Space Domain Expertise Experience: 5 years of progressive systems engineering experience in satellite operations, space domain awareness, or space traffic management environments. Technical Foundations: A deep, demonstrable understanding of astrodynamics, orbital mechanics, covariance realism, maneuver planning, and conjunction assessment (CA) methodologies. Data Competency: Familiarity with the ingestion, processing, and dissemination of Space Situational Awareness (SSA) data products. Background Summary: The Office of Space Commerce (OSC), is developing the Traffic Coordination System for Space (TraCSS) to fulfill Space Policy Directive 3 (SPD-3). SPD-3 instructed relevant US government agencies to begin re-assigning many aspects of space traffic management (STM) and space traffic coordination (STC) serving non-military US space operators. OSC, under the US Department of Commerce (DOC), was identified to lead many of these efforts as part of a 'whole of government' approach. TraCSS is the Office of Space Commerce's cloud-based enterprise solution for ingesting, archiving, processing, and disseminating Space Situational Awareness (SSA) data and products. It will provide conjunction analysis and warning services to commercial satellite owner/operators to foster economic growth and technological advancement of the US commercial space industry. The system will store data from the Department of Defense (DoD), NOAA, commercial SSA data providers, commercial and civil satellite Owner/Operators (O/O), and select international civil partners. The TraCSS system will operate 24 hours per day, 7 days a week. Work Environment and Physical Demands: Work in a typical onsite government office building full time (5-days per week) with situational telework approved as needed. Occasional deadlines or operational conditions may require non-standard/longer hours or additional onsite support. Infrequent travel required. Place of performance will be the NOAA Satellite Operations Facility located on the Suitland Federal Center in Suitland MD. We invest in the lives of our employees, both in and out of the workplace, by providing competitive pay and benefits packages. Benefits offered may include health care, dental, vision, life insurance; 401(k); education assistance; paid time off including PTO, holidays, and any other paid leave required by law. The salary offered will depend on several factors including, but not limited to, relevant experience, skills, education, geographic location, internal equity, business needs, and other factors permitted by law. Posted pay ranges are a general guideline only and are not a guarantee of compensation or salary. EEO Statement ASRC Federal and its Subsidiaries are Equal Opportunity employers. All qualified applicants will receive consideration for employment without regard to race, gender, color, age, sexual orientation, gender identification, national origin, religion, marital status, ancestry, citizenship, disability, protected veteran status, or any other factor prohibited by applicable law.
09/23/2026
Full time
ASRC Federal is a leading government contractor furthering missions in space, public health and defense. As an Alaska Native owned corporation, our work helps secure an enduring future for our shareholders. Join our team and discover why we are a top veteran employer and Certified Great Place to Work ASRC Federal Data Networks Corporation (DNC) is seeking a senior, versatile, Systems Engineer to serve as the critical technical integrator within the Traffic Coordination System for Space (TraCSS) Scaled Agile Framework (SAFe) ecosystem. Operating at the Release and Requirements Management level, this role bridges the gap between high-level government strategy and multi-vendor execution. Our contract team supports the Office of Space Commerce, Space Operations Division, TraCSS Engineering Branch located in Suitland MD. The successful candidate will take strategic direction from the Government Business Owner, Program Management, and Product Owners to ensure systemic technical integrity across the TraCSS enterprise. Acting as a collaborative technical leader, you will work hand-in-hand with the System Integrator's Release Train Engineer (RTE) and coordinate directly with the System Integrator and Presentation Layer Scrum Masters to clear program-level technical roadblocks, refine backlogs, and maintain a robust architectural runway. Key Responsibilities: SAFe Technical Leadership: Own the technical "big picture" for system integration across the Agile Release Train (ART). Partner with the System Integrator's RTE to ensure multi-vendor scrum teams remain aligned with the broader product vision. Cross-Vendor Integration & Synchronization: Coordinate directly with the Presentation Layer and System Integrator contractors, satellite Owner/Operators (O/Os), and commercial vendors to identify, analyze, and resolve complex engineering and integration issues. Requirements & Backlog Refinement: Collaborate closely with the TraCSS Requirements Engineer to translate high-level stakeholder needs into technical clarity. Assist Product Owners in Epic/Story creation, technical dependency mapping, and backlog refinement. Interagency & Industry Liaison: Represent the Systems Engineering integrated program team in technical discussions and meetings with commercial space situational awareness providers, civil operators, and external federal agencies (including NASA and industry partners) to track studies and align data standards. Mandatory Qualifications & Criteria: 1. Agile & Framework Proficiency SAFe Ecosystems : Proven experience operating as a System Engineer, Solution Architect, or Product Owner within a Scaled Agile Framework (SAFe) environment or on a large-scale Agile Release Train (ART). Interface Coordination : Documented success bridging gaps between product management levels and technical Scrum Masters/Developers in a multi-contractor configuration. Tooling : Proficiency with the Atlassian Suite (Jira, Jira Service Management, Confluence) for tracking requirements, dependencies, and program risks. 2. Compliance & Education Education : Bachelor's degree in Aerospace Engineering, Physics, Mathematics, Computer Science, or a related technical field. Background Processing : Must successfully pass NOAA background screening requirements and sign a mandatory Office of Space Commerce Non-Disclosure Agreement (NDA). Citizenship : US Citizenship required. Highly Desired 1. Space Domain Expertise Experience: 5 years of progressive systems engineering experience in satellite operations, space domain awareness, or space traffic management environments. Technical Foundations: A deep, demonstrable understanding of astrodynamics, orbital mechanics, covariance realism, maneuver planning, and conjunction assessment (CA) methodologies. Data Competency: Familiarity with the ingestion, processing, and dissemination of Space Situational Awareness (SSA) data products. Background Summary: The Office of Space Commerce (OSC), is developing the Traffic Coordination System for Space (TraCSS) to fulfill Space Policy Directive 3 (SPD-3). SPD-3 instructed relevant US government agencies to begin re-assigning many aspects of space traffic management (STM) and space traffic coordination (STC) serving non-military US space operators. OSC, under the US Department of Commerce (DOC), was identified to lead many of these efforts as part of a 'whole of government' approach. TraCSS is the Office of Space Commerce's cloud-based enterprise solution for ingesting, archiving, processing, and disseminating Space Situational Awareness (SSA) data and products. It will provide conjunction analysis and warning services to commercial satellite owner/operators to foster economic growth and technological advancement of the US commercial space industry. The system will store data from the Department of Defense (DoD), NOAA, commercial SSA data providers, commercial and civil satellite Owner/Operators (O/O), and select international civil partners. The TraCSS system will operate 24 hours per day, 7 days a week. Work Environment and Physical Demands: Work in a typical onsite government office building full time (5-days per week) with situational telework approved as needed. Occasional deadlines or operational conditions may require non-standard/longer hours or additional onsite support. Infrequent travel required. Place of performance will be the NOAA Satellite Operations Facility located on the Suitland Federal Center in Suitland MD. We invest in the lives of our employees, both in and out of the workplace, by providing competitive pay and benefits packages. Benefits offered may include health care, dental, vision, life insurance; 401(k); education assistance; paid time off including PTO, holidays, and any other paid leave required by law. The salary offered will depend on several factors including, but not limited to, relevant experience, skills, education, geographic location, internal equity, business needs, and other factors permitted by law. Posted pay ranges are a general guideline only and are not a guarantee of compensation or salary. EEO Statement ASRC Federal and its Subsidiaries are Equal Opportunity employers. All qualified applicants will receive consideration for employment without regard to race, gender, color, age, sexual orientation, gender identification, national origin, religion, marital status, ancestry, citizenship, disability, protected veteran status, or any other factor prohibited by applicable law.
About Hightouch Hightouch is an Agentic Marketing Platform powered by the industry-leading Composable CDP. With complete brand context, customer data, and performance history in one place, every marketer finally has the power to build and ship end-to-end campaigns themselves. Teams move faster, stay on brand, and get AI marketing that actually works. Founded in 2019 and headquartered in San Francisco, Hightouch enables marketing teams to analyze performance, brainstorm ideas, and generate creative at a speed and quality that wasn't previously possible. Named a Leader in the 2026 Gartner Magic Quadrant for Customer Data Platforms, Hightouch is trusted by leading enterprises like Domino's, Spotify, Aritzia, Ramp, and PetSmart. At Hightouch, our mission is to help our customers leverage data and AI to grow their businesses. The team is ambitious, impact-driven, efficient - and we believe humility, kindness, and compassion are essential to our success. If you're energized by velocity, obsessed with raising the bar, and want to build alongside people who care deeply about each other and our customers, we'd love to meet you. About the Role We are looking for a developer productivity engineer to take responsibility for our monorepo and the "path to production" for over 50 engineers pushing over 75 commits a day with continuous deployment to production. This presents an exciting challenge where you can apply your expertise in helping teams ship fast and safe to meaningfully improve the productivity of our engineering team. This role also provides a unique opportunity to work on a multi-cloud and multi-region infrastructure that supports a global customer base. Our monorepo is primarily Javascript/Typescript with some Go and Python. While the primary responsibility of this job is developer productivity, to fit in with the team and be productive in our codebase generally, strong computer science and development fundamentals will be required. As part of your onboarding, you'll spend significant time writing and delivering features to gain empathy for the current dev flow. We believe in enabling our engineers to do their best work for our customers by giving them extremely high levels of ownership and autonomy. This comes in different forms: you will own and deliver projects from start to finish, you will work directly with customers to solve their hardest scaling problems, and you will have a lot of influence over what we work on as a team and company. Some of the problems we'll be working on include: Own the build : You'll be the single threaded owner for the build/test/deploy of our software and how each team fits into it Monorepo productivity: Detangle our build/test/deploy patterns so teams can move fast and not block each other. Investigate/implement a tool like turbo repo to speed builds and separate concerns Drive excellence in testing: We need to improve our top-down and team-level views into test coverage + support an ever growing matrix of data sources, data destinations, and enrichment patterns. An excellent engineer in this role will be able to "hold up a mirror" to our dev teams, helping each team understand where they have gaps Multi-Region and Multi-Cloud: Supporting our multi-region and multi-cloud backend, including extending it to launch Hightouch on in new regions to support data residency requirements of our global customer base Operational excellence: Support an increase in our ability to catch issues before they reach production and respond quickly if they do We are looking for talented, intellectually curious, and motivated individuals who are interested in tackling the problems above. This is a senior role, but we focus on impact and potential for growth more than years of experience. The salary range for this position is $180,000 - $400,000 USD per year, which is location independent in accordance with our remote-first policy. We also offer meaningful equity compensation. About You You are an engineer with a passion for solving hard technical problems that generate real value for customers. You're motivated by high ownership and are comfortable in a fast-paced, startup environment. You've driven significant improvement in the productivity of a 50+ person development team by making high-leverage changes to their build/test/deploy processes. While you likely have some infra chops, this is not an infra role: you have strong development fundamentals and are comfortable driving framework-level improvements across multiple teams. This could take several forms: You've been a member of a developer-productivity team/function for an outstanding development team You've organically become "the build person" at a fast-growing startup and helped bring order to chaos You've personally delivered code that has improved cross-team metrics related to developer productivity (e.g., DORA metrics, coverage, etc) In addition to any of the above, we're looking for someone who can dive into feature code, including complicated backend code, as needed. Interview Process Our goal with the interview process is to balance speed with giving both parties opportunities to assess whether there is a strong mutual fit. We will ask you questions, but we want you to ask us questions! Our technical interviews focus on how you design systems because we believe this is the best way for us to see how you work and for you to see how we collaborate. We don't ask you to write code to solve technical brainteasers that don't appear in your day to day job. Apply: Curl on port 13784 and have followed those instructions before applying. It'll only take a minute! Recruiter Screen: Introductory call with our recruiting team to get to know each other and see if the role could be a good mutual fit System Design Screen: Designing a data processing feature end-to-end. Developer Productivity Skills Interview: We'll dive into your experience making teams productive through a discussion of past projects and tooling you've used. Hiring Manager Interview: Chat with hiring manager about past experiences and future operating preferences to assess fit on company values and operating principles. System Design Interview: Work with the interviewer to architect a system at a conceptual level. The problem will be at a pretty high level - and have both product and customer requirements as well as technical. We have limited inbound applications to one application per candidate. You will be auto-rejected if you apply to multiple roles. Please only apply to the position you are most qualified for. E-Verify Statement Hightouch participates in E-Verify. After you join the team, we'll verify your eligibility to work in the U.S. by submitting information from your Form I-9 to the Social Security Administration and, if needed, the Department of Homeland Security. This process happens post-hire only - we never use E-Verify to pre-screen applicants. E-Verify Notice E-Verify Notice (Spanish) Right to Work Notice Right to Work Notice (Spanish)
09/23/2026
Full time
About Hightouch Hightouch is an Agentic Marketing Platform powered by the industry-leading Composable CDP. With complete brand context, customer data, and performance history in one place, every marketer finally has the power to build and ship end-to-end campaigns themselves. Teams move faster, stay on brand, and get AI marketing that actually works. Founded in 2019 and headquartered in San Francisco, Hightouch enables marketing teams to analyze performance, brainstorm ideas, and generate creative at a speed and quality that wasn't previously possible. Named a Leader in the 2026 Gartner Magic Quadrant for Customer Data Platforms, Hightouch is trusted by leading enterprises like Domino's, Spotify, Aritzia, Ramp, and PetSmart. At Hightouch, our mission is to help our customers leverage data and AI to grow their businesses. The team is ambitious, impact-driven, efficient - and we believe humility, kindness, and compassion are essential to our success. If you're energized by velocity, obsessed with raising the bar, and want to build alongside people who care deeply about each other and our customers, we'd love to meet you. About the Role We are looking for a developer productivity engineer to take responsibility for our monorepo and the "path to production" for over 50 engineers pushing over 75 commits a day with continuous deployment to production. This presents an exciting challenge where you can apply your expertise in helping teams ship fast and safe to meaningfully improve the productivity of our engineering team. This role also provides a unique opportunity to work on a multi-cloud and multi-region infrastructure that supports a global customer base. Our monorepo is primarily Javascript/Typescript with some Go and Python. While the primary responsibility of this job is developer productivity, to fit in with the team and be productive in our codebase generally, strong computer science and development fundamentals will be required. As part of your onboarding, you'll spend significant time writing and delivering features to gain empathy for the current dev flow. We believe in enabling our engineers to do their best work for our customers by giving them extremely high levels of ownership and autonomy. This comes in different forms: you will own and deliver projects from start to finish, you will work directly with customers to solve their hardest scaling problems, and you will have a lot of influence over what we work on as a team and company. Some of the problems we'll be working on include: Own the build : You'll be the single threaded owner for the build/test/deploy of our software and how each team fits into it Monorepo productivity: Detangle our build/test/deploy patterns so teams can move fast and not block each other. Investigate/implement a tool like turbo repo to speed builds and separate concerns Drive excellence in testing: We need to improve our top-down and team-level views into test coverage + support an ever growing matrix of data sources, data destinations, and enrichment patterns. An excellent engineer in this role will be able to "hold up a mirror" to our dev teams, helping each team understand where they have gaps Multi-Region and Multi-Cloud: Supporting our multi-region and multi-cloud backend, including extending it to launch Hightouch on in new regions to support data residency requirements of our global customer base Operational excellence: Support an increase in our ability to catch issues before they reach production and respond quickly if they do We are looking for talented, intellectually curious, and motivated individuals who are interested in tackling the problems above. This is a senior role, but we focus on impact and potential for growth more than years of experience. The salary range for this position is $180,000 - $400,000 USD per year, which is location independent in accordance with our remote-first policy. We also offer meaningful equity compensation. About You You are an engineer with a passion for solving hard technical problems that generate real value for customers. You're motivated by high ownership and are comfortable in a fast-paced, startup environment. You've driven significant improvement in the productivity of a 50+ person development team by making high-leverage changes to their build/test/deploy processes. While you likely have some infra chops, this is not an infra role: you have strong development fundamentals and are comfortable driving framework-level improvements across multiple teams. This could take several forms: You've been a member of a developer-productivity team/function for an outstanding development team You've organically become "the build person" at a fast-growing startup and helped bring order to chaos You've personally delivered code that has improved cross-team metrics related to developer productivity (e.g., DORA metrics, coverage, etc) In addition to any of the above, we're looking for someone who can dive into feature code, including complicated backend code, as needed. Interview Process Our goal with the interview process is to balance speed with giving both parties opportunities to assess whether there is a strong mutual fit. We will ask you questions, but we want you to ask us questions! Our technical interviews focus on how you design systems because we believe this is the best way for us to see how you work and for you to see how we collaborate. We don't ask you to write code to solve technical brainteasers that don't appear in your day to day job. Apply: Curl on port 13784 and have followed those instructions before applying. It'll only take a minute! Recruiter Screen: Introductory call with our recruiting team to get to know each other and see if the role could be a good mutual fit System Design Screen: Designing a data processing feature end-to-end. Developer Productivity Skills Interview: We'll dive into your experience making teams productive through a discussion of past projects and tooling you've used. Hiring Manager Interview: Chat with hiring manager about past experiences and future operating preferences to assess fit on company values and operating principles. System Design Interview: Work with the interviewer to architect a system at a conceptual level. The problem will be at a pretty high level - and have both product and customer requirements as well as technical. We have limited inbound applications to one application per candidate. You will be auto-rejected if you apply to multiple roles. Please only apply to the position you are most qualified for. E-Verify Statement Hightouch participates in E-Verify. After you join the team, we'll verify your eligibility to work in the U.S. by submitting information from your Form I-9 to the Social Security Administration and, if needed, the Department of Homeland Security. This process happens post-hire only - we never use E-Verify to pre-screen applicants. E-Verify Notice E-Verify Notice (Spanish) Right to Work Notice Right to Work Notice (Spanish)
About Trackonomy Trackonomy is pioneering a transformative network of interconnected objects. Our goal is to bring inanimate objects to life, enabling them to communicate, think, and interact in real time. We are building the operating system for the connected world-one object and sensor at a time. Our customers span logistics, industrial, utilities, healthcare, and government sectors, using our solutions for predictive maintenance, workflow optimization, asset protection, safety, security, and environmental monitoring. Backed by leading investors including Kleiner Perkins and 8VC, Trackonomy is one of Silicon Valley's fastest-growing IoT companies. Our executive leadership team has worked together for more than two decades, holding key leadership positions at companies including Flextronics, Heptagon, GT-Nexus, and Digital Motors Corporation. Together, they have engineered breakthrough technology that is delivering extraordinary results for customers around the world. Are you ready to be part of the next chapter in Silicon Valley's hyper-growth IoT success story? The Role You will be a core engineer on our early-stage team. We work on everything from machine learning and security to high-performance computing, IoT devices, and dynamic web applications. Don't be surprised if you have the opportunity to touch nearly every system while working here, often collaborating with teammates on critical initiatives. Candidates must be comfortable working with microcontrollers and low-level hardware control in a test-driven development environment. The ideal candidate will have experience working with wireless communications modules, IoT technologies, RF protocols, and strong troubleshooting and prototyping skills. Experience Bachelor's or Master's degree in Electronics Engineering At least 6+ years of embedded software development with emphasis on C/C++ Background in real-time embedded systems Extensive work in firmware development, testing, and system-level bring-up and debugging Skilled in bench modifications and rapid prototyping of hardware/firmware solutions Strong interpersonal, organizational, and communication skills Effective collaborator who shares knowledge, learns from others, and supports cross-functional teams Self-starter with strong motivation and ownership mentality Preferred Skills Background in IoT systems and wireless/wired communication protocols, including BLE, LoRa, and LoRaWAN Skilled in firmware development on ARM-based microcontrollers and processors; familiarity with Nordic devices is a plus Knowledge of cellular protocols such as LTE, LTE-M/NB-IoT, and industrial wireless electronics Hands-on work with cellular modems and GPS/GNSS systems Ability to develop device drivers for sensors and communication modems Strong understanding of subsystem interfaces such as UART, I2C, SPI, and other standard chip-level protocols Proficiency in low-power embedded development, including performing power profiling on target devices Knowledge of firmware development best practices including testing, documentation, debugging, and code review Understanding of PCB design using schematic capture and layout tools (Eagle/Altium preferred) Capability to design bootloaders and implement firmware-over-the-air updates Familiarity with scripting languages (Python preferred) Ability to read and interpret complex electrical schematics Strong foundation in microcontrollers and embedded peripheral driver development Knowledge of embedded networking protocols such as RSTP, PTP, LLDP, and UDP/TCP is a plus Why Trackonomy? Meaningful Impact Trackonomy is dedicated to building technology that solves real-world problems. From fire prevention and environmental monitoring to safety, security, and operational efficiency, our innovations help organizations protect assets, improve outcomes, and save lives. Growth & Development At Trackonomy, your growth is driven by your capabilities, passion, ownership, and results. As part of an early-stage, high-growth company, you will have opportunities to take on diverse responsibilities, learn directly from experienced leaders, and expand your role as your impact grows. We believe talent and execution matter more than hierarchy. Team members are encouraged to explore new challenges, collaborate across functions, and continuously develop new skills. Ownership & Collaboration Our culture isn't something you join-it's something you help build. Every employee plays a meaningful role in our success and contributes to shaping our products, processes, and future. We value collaboration, knowledge sharing, accountability, and a strong sense of ownership. Benefits & Rewards Trackonomy understands that personal wellness is critical to a happy, healthy, and productive work environment. We offer: Platinum-level health benefits Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) Commuter benefits Employee Assistance Program (EAP) 401(k) plan Pre-IPO equity opportunity Regular performance reviews and feedback Learning and development opportunities for both individual contributors and aspiring leaders Equal Opportunity Employer Trackonomy Systems is proud to provide equal employment opportunities to all individuals regardless of race, color, religion, national origin, ancestry, physical or mental disability, sex, gender, gender identity, gender expression, sexual orientation, age, medical condition, genetic information, marital or registered domestic partnership status, military or veteran status, or any other characteristic protected by applicable law. We strive to provide a stellar experience throughout the application process and ensure all applicants receive fair consideration based solely on merit and business needs. The salary range for this role is $140,000 to $200,000, plus bonuses and Pre-IPO equity. It is uncommon for anyone to be hired at or near the top of the range. Final compensation is based on several factors, including skills, experience, training, business needs, cultural fit, level, and location. Trackonomy Systems is dedicated to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need assistance or accommodation during the interview process, please contact . When you apply to a job on this site, you acknowledge and agree that the personal data contained in your application will be collected and processed by Trackonomy Systems, Inc. and/or one of its subsidiaries ("Trackonomy") in accordance with our Applicant Privacy Notice. If you have any questions about our privacy practices, please contact .
09/23/2026
Full time
About Trackonomy Trackonomy is pioneering a transformative network of interconnected objects. Our goal is to bring inanimate objects to life, enabling them to communicate, think, and interact in real time. We are building the operating system for the connected world-one object and sensor at a time. Our customers span logistics, industrial, utilities, healthcare, and government sectors, using our solutions for predictive maintenance, workflow optimization, asset protection, safety, security, and environmental monitoring. Backed by leading investors including Kleiner Perkins and 8VC, Trackonomy is one of Silicon Valley's fastest-growing IoT companies. Our executive leadership team has worked together for more than two decades, holding key leadership positions at companies including Flextronics, Heptagon, GT-Nexus, and Digital Motors Corporation. Together, they have engineered breakthrough technology that is delivering extraordinary results for customers around the world. Are you ready to be part of the next chapter in Silicon Valley's hyper-growth IoT success story? The Role You will be a core engineer on our early-stage team. We work on everything from machine learning and security to high-performance computing, IoT devices, and dynamic web applications. Don't be surprised if you have the opportunity to touch nearly every system while working here, often collaborating with teammates on critical initiatives. Candidates must be comfortable working with microcontrollers and low-level hardware control in a test-driven development environment. The ideal candidate will have experience working with wireless communications modules, IoT technologies, RF protocols, and strong troubleshooting and prototyping skills. Experience Bachelor's or Master's degree in Electronics Engineering At least 6+ years of embedded software development with emphasis on C/C++ Background in real-time embedded systems Extensive work in firmware development, testing, and system-level bring-up and debugging Skilled in bench modifications and rapid prototyping of hardware/firmware solutions Strong interpersonal, organizational, and communication skills Effective collaborator who shares knowledge, learns from others, and supports cross-functional teams Self-starter with strong motivation and ownership mentality Preferred Skills Background in IoT systems and wireless/wired communication protocols, including BLE, LoRa, and LoRaWAN Skilled in firmware development on ARM-based microcontrollers and processors; familiarity with Nordic devices is a plus Knowledge of cellular protocols such as LTE, LTE-M/NB-IoT, and industrial wireless electronics Hands-on work with cellular modems and GPS/GNSS systems Ability to develop device drivers for sensors and communication modems Strong understanding of subsystem interfaces such as UART, I2C, SPI, and other standard chip-level protocols Proficiency in low-power embedded development, including performing power profiling on target devices Knowledge of firmware development best practices including testing, documentation, debugging, and code review Understanding of PCB design using schematic capture and layout tools (Eagle/Altium preferred) Capability to design bootloaders and implement firmware-over-the-air updates Familiarity with scripting languages (Python preferred) Ability to read and interpret complex electrical schematics Strong foundation in microcontrollers and embedded peripheral driver development Knowledge of embedded networking protocols such as RSTP, PTP, LLDP, and UDP/TCP is a plus Why Trackonomy? Meaningful Impact Trackonomy is dedicated to building technology that solves real-world problems. From fire prevention and environmental monitoring to safety, security, and operational efficiency, our innovations help organizations protect assets, improve outcomes, and save lives. Growth & Development At Trackonomy, your growth is driven by your capabilities, passion, ownership, and results. As part of an early-stage, high-growth company, you will have opportunities to take on diverse responsibilities, learn directly from experienced leaders, and expand your role as your impact grows. We believe talent and execution matter more than hierarchy. Team members are encouraged to explore new challenges, collaborate across functions, and continuously develop new skills. Ownership & Collaboration Our culture isn't something you join-it's something you help build. Every employee plays a meaningful role in our success and contributes to shaping our products, processes, and future. We value collaboration, knowledge sharing, accountability, and a strong sense of ownership. Benefits & Rewards Trackonomy understands that personal wellness is critical to a happy, healthy, and productive work environment. We offer: Platinum-level health benefits Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) Commuter benefits Employee Assistance Program (EAP) 401(k) plan Pre-IPO equity opportunity Regular performance reviews and feedback Learning and development opportunities for both individual contributors and aspiring leaders Equal Opportunity Employer Trackonomy Systems is proud to provide equal employment opportunities to all individuals regardless of race, color, religion, national origin, ancestry, physical or mental disability, sex, gender, gender identity, gender expression, sexual orientation, age, medical condition, genetic information, marital or registered domestic partnership status, military or veteran status, or any other characteristic protected by applicable law. We strive to provide a stellar experience throughout the application process and ensure all applicants receive fair consideration based solely on merit and business needs. The salary range for this role is $140,000 to $200,000, plus bonuses and Pre-IPO equity. It is uncommon for anyone to be hired at or near the top of the range. Final compensation is based on several factors, including skills, experience, training, business needs, cultural fit, level, and location. Trackonomy Systems is dedicated to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need assistance or accommodation during the interview process, please contact . When you apply to a job on this site, you acknowledge and agree that the personal data contained in your application will be collected and processed by Trackonomy Systems, Inc. and/or one of its subsidiaries ("Trackonomy") in accordance with our Applicant Privacy Notice. If you have any questions about our privacy practices, please contact .
The Role The Global Operations Technology team at Schonfeld is tasked with creating a platform that serves as the golden data source for post trade operational data - transaction capture, near real-time positions, start-of-day position, end-of-day position, trades, executions and P&L. The team is responsible for Drop copy and FIX integrations for post trade across our multi-manager multi-strategy platform. Building and managing positions across a wide array of asset classes including listed and OTC products. Building and maintaining all systems related to the fund's intraday and end of day trade reporting facilities across all investments to our Prime Brokers, Fund Primary and Shadow books of record systems. What you'll do Senior Software Engineers at Schonfeld take on a wide range of responsibilities and challenges. They design, develop, deploy and support existing applications as well as new features and platforms. They run projects and lead initiatives. You will participate in an Agile framework, continuously improving and expanding platform capabilities based on rapidly changing business requirements as well as providing Level 3 support across the globe. You will contribute to a highly supportable, future-ready, performant code base. You will work with a modern tech stack - building microservices that deploy to the cloud (AWS) using containers (Docker) orchestrated by Kubernetes with Helm. The ideal candidate will have a proven ability to solve problems and an interest in continuous improvement and learning. What you'll bring What you need: At minimum, seven years of experience in financial services, working on applications supporting the trade lifecycle At minimum, seven years of Java development Expertise with design and architecture (microservices experience preferred) Experience leading strategic technical initiatives from start to finish Experience with relational databases (PostgreSQL preferred) Experience with building and consuming from REST APIs Experience with messaging systems (Kafka preferred) Excellent communication skills, both written and verbal Experience mentoring junior developers Strong ownership experience and a track record of delivering results We'd love if you had: Experience supporting listed and OTC instruments typical in a Global Macro strategy (Fixed Income, Credit, Rates, Equities, Futures) Experience developing and supporting stateful applications build on Apache Flink Experience and knowledge of: Cloud services such as AWS Containerization and orchestration tools such as Docker and Kubernetes with Helm. DevOps methodologies such as CI/CD and build automation Who we are Schonfeld Strategic Advisors is a global multi-strategy, multi-manager investment platform that harnesses the transformative power of people to perform in all market environments. Our dynamic culture inspires better outcomes for our team, our investors, and our partners. We aim to consistently deliver risk-adjusted returns, with people driving performance. We specialize in four core strategies: Quantitative Trading, Fundamental Equity, Tactical Trading, and Discretionary Macro & Fixed Income. We capitalize on inefficiencies and opportunities within the markets, drawing from a significant investment in proprietary technology, infrastructure, and risk analytics. We invest through internal portfolio managers and external partner funds, pursuing alignment among investors, investment professionals, and the firm. Our footprint spans 7 countries and 19 offices. Our Culture Talent is our strategy. We believe our success is because of our people, so putting our talent above all else is our top priority. We are teamwork-oriented, collaborative, and encourage ideas-at all levels-to be shared. As an organization committed to investing in our people, we provide learning & educational offerings and opportunities to make an impact. We foster a sense of belonging among all of our employees with Diversity, Equity and Inclusion at the forefront of this mission. Our employees value diversity across identity, thought, people and perspective which serves as the foundation of our culture. As a firm, we are committed to creating a hiring process that is fair, welcoming and supportive. The annual base pay for this role is expected to be between $180,000 and $250,000 which will be prorated based on start and end date. The expected base pay range is based on information at the time this post was generated. Actual compensation for the successful candidate will be determined based on a variety of factors such as skills, qualifications and experience and level of education. Our Culture The firm's ethos is embedded in our people. 'Talent is our strategy' is our mantra and drives how we approach all initiatives at the firm. We believe our success is because of our people, so putting our talent above all else is our top priority. Schonfeld strives to create an environment where our people can thrive. We foster a teamwork-oriented, collaborative environment where ideas at any level are encouraged and shared. The development and advancement of our talent is honed through interactions with each other, learning & educational offerings, and through opportunities to make impactful contributions. At Schonfeld, we strive to cultivate a sense of belonging throughout all of our employees with Diversity, Equity and Inclusion at the forefront of this mission. As a firm we are committed to creating a hiring process which is not only fair, but also welcoming and supportive. On a daily basis, our employees welcome diversity across identity, thought, people and views which serves as the foundation of our culture and success. You can learn more about our DEI initiatives here - Who we are Schonfeld Strategic Advisors is a multi-manager platform that invests its capital with Internal and Partner portfolio managers, primarily on an exclusive or semi-exclusive basis, across four trading strategies; quantitative, fundamental equity, tactical trading and discretionary macro & fixed income. We have created a unique structure to provide global portfolio managers with autonomy, flexibility and support to best enable them to maximize the value of their businesses. Over the last 30 years, Schonfeld has successfully capitalized on inefficiencies and opportunities within the markets. We have developed and invested heavily in proprietary technology, infrastructure and risk analytics and continue to capitalize on new opportunities. In 2021 we launched our newest strategy, discretionary macro & fixed income as part of the continual growth of Schonfeld's investible universe. Our portfolio exposure has expanded across the Americas, Europe and Asia as well as multiple asset classes and products . The base pay for this role is expected to be between $90,000 and $150,000. The expected base pay range is based on information at the time this post was generated. This role may also be eligible for other forms of compensation such as a performance bonus and a competitive benefits package. Actual compensation for the successful candidate will be determined based on a variety of factors such as skills, qualifications, and experience.
09/23/2026
Full time
The Role The Global Operations Technology team at Schonfeld is tasked with creating a platform that serves as the golden data source for post trade operational data - transaction capture, near real-time positions, start-of-day position, end-of-day position, trades, executions and P&L. The team is responsible for Drop copy and FIX integrations for post trade across our multi-manager multi-strategy platform. Building and managing positions across a wide array of asset classes including listed and OTC products. Building and maintaining all systems related to the fund's intraday and end of day trade reporting facilities across all investments to our Prime Brokers, Fund Primary and Shadow books of record systems. What you'll do Senior Software Engineers at Schonfeld take on a wide range of responsibilities and challenges. They design, develop, deploy and support existing applications as well as new features and platforms. They run projects and lead initiatives. You will participate in an Agile framework, continuously improving and expanding platform capabilities based on rapidly changing business requirements as well as providing Level 3 support across the globe. You will contribute to a highly supportable, future-ready, performant code base. You will work with a modern tech stack - building microservices that deploy to the cloud (AWS) using containers (Docker) orchestrated by Kubernetes with Helm. The ideal candidate will have a proven ability to solve problems and an interest in continuous improvement and learning. What you'll bring What you need: At minimum, seven years of experience in financial services, working on applications supporting the trade lifecycle At minimum, seven years of Java development Expertise with design and architecture (microservices experience preferred) Experience leading strategic technical initiatives from start to finish Experience with relational databases (PostgreSQL preferred) Experience with building and consuming from REST APIs Experience with messaging systems (Kafka preferred) Excellent communication skills, both written and verbal Experience mentoring junior developers Strong ownership experience and a track record of delivering results We'd love if you had: Experience supporting listed and OTC instruments typical in a Global Macro strategy (Fixed Income, Credit, Rates, Equities, Futures) Experience developing and supporting stateful applications build on Apache Flink Experience and knowledge of: Cloud services such as AWS Containerization and orchestration tools such as Docker and Kubernetes with Helm. DevOps methodologies such as CI/CD and build automation Who we are Schonfeld Strategic Advisors is a global multi-strategy, multi-manager investment platform that harnesses the transformative power of people to perform in all market environments. Our dynamic culture inspires better outcomes for our team, our investors, and our partners. We aim to consistently deliver risk-adjusted returns, with people driving performance. We specialize in four core strategies: Quantitative Trading, Fundamental Equity, Tactical Trading, and Discretionary Macro & Fixed Income. We capitalize on inefficiencies and opportunities within the markets, drawing from a significant investment in proprietary technology, infrastructure, and risk analytics. We invest through internal portfolio managers and external partner funds, pursuing alignment among investors, investment professionals, and the firm. Our footprint spans 7 countries and 19 offices. Our Culture Talent is our strategy. We believe our success is because of our people, so putting our talent above all else is our top priority. We are teamwork-oriented, collaborative, and encourage ideas-at all levels-to be shared. As an organization committed to investing in our people, we provide learning & educational offerings and opportunities to make an impact. We foster a sense of belonging among all of our employees with Diversity, Equity and Inclusion at the forefront of this mission. Our employees value diversity across identity, thought, people and perspective which serves as the foundation of our culture. As a firm, we are committed to creating a hiring process that is fair, welcoming and supportive. The annual base pay for this role is expected to be between $180,000 and $250,000 which will be prorated based on start and end date. The expected base pay range is based on information at the time this post was generated. Actual compensation for the successful candidate will be determined based on a variety of factors such as skills, qualifications and experience and level of education. Our Culture The firm's ethos is embedded in our people. 'Talent is our strategy' is our mantra and drives how we approach all initiatives at the firm. We believe our success is because of our people, so putting our talent above all else is our top priority. Schonfeld strives to create an environment where our people can thrive. We foster a teamwork-oriented, collaborative environment where ideas at any level are encouraged and shared. The development and advancement of our talent is honed through interactions with each other, learning & educational offerings, and through opportunities to make impactful contributions. At Schonfeld, we strive to cultivate a sense of belonging throughout all of our employees with Diversity, Equity and Inclusion at the forefront of this mission. As a firm we are committed to creating a hiring process which is not only fair, but also welcoming and supportive. On a daily basis, our employees welcome diversity across identity, thought, people and views which serves as the foundation of our culture and success. You can learn more about our DEI initiatives here - Who we are Schonfeld Strategic Advisors is a multi-manager platform that invests its capital with Internal and Partner portfolio managers, primarily on an exclusive or semi-exclusive basis, across four trading strategies; quantitative, fundamental equity, tactical trading and discretionary macro & fixed income. We have created a unique structure to provide global portfolio managers with autonomy, flexibility and support to best enable them to maximize the value of their businesses. Over the last 30 years, Schonfeld has successfully capitalized on inefficiencies and opportunities within the markets. We have developed and invested heavily in proprietary technology, infrastructure and risk analytics and continue to capitalize on new opportunities. In 2021 we launched our newest strategy, discretionary macro & fixed income as part of the continual growth of Schonfeld's investible universe. Our portfolio exposure has expanded across the Americas, Europe and Asia as well as multiple asset classes and products . The base pay for this role is expected to be between $90,000 and $150,000. The expected base pay range is based on information at the time this post was generated. This role may also be eligible for other forms of compensation such as a performance bonus and a competitive benefits package. Actual compensation for the successful candidate will be determined based on a variety of factors such as skills, qualifications, and experience.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Over the last 20 years, Ares' success has been driven by our people and our culture. Today, our team is guided by our core values - Collaborative, Responsible, Entrepreneurial, Self-Aware, Trustworthy - and our purpose to be a catalyst for shared prosperity and a better future. Through our recruitment, career development and employee-focused programming, we are committed to fostering a welcoming and inclusive work environment where high-performance talent of diverse backgrounds, experiences, and perspectives can build careers within this exciting and growing industry. Job Description Ares is seeking a Legal Engineering Lead (Associate Vice President) to join Corporate Technology, reporting to the Head of Legal and Compliance Technology. This role will lead engineering delivery, platform health, and ongoing evolution of technology solutions supporting Ares' Legal and related risk and third-party governance workflows-driving scalable, secure, and audit-ready capabilities that improve operational efficiency and control effectiveness. The ideal candidate is a hands-on engineering leader with deep familiarity across Legal/GC processes and related risk workflows, ideally from a large alternative asset manager or similarly complex regulated firm (large law firm experience supporting sophisticated legal operations is also valued). This leader will bring strong technical fundamentals (e.g., SQL and at least one coding language), experience with integrations and data platforms, and the ability to partner closely with business stakeholders and vendors to deliver measurable outcomes. This role requires experience implementing and operating modern legal and risk platforms, including matter management and e-billing (e.g., Simple Legal / Passport / Apperio or similar), Legal Entity Management (e.g., GEMS), Contract Lifecycle Management (CLM) solutions (e.g., Icertis, Ironclad), and third-party risk management (TPRM) and intake tooling (e.g., Zip and comparable platforms). RESPONSIBILITIES: • Lead engineering delivery and operational ownership for the Legal and technology portfolio, ensuring solutions are scalable, resilient, secure, and audit-ready. • Serve as the technical lead for platform strategy, implementation, and support across Legal systems, including matter management/e-billing, legal entity management, CLM, and third-party intake/TPRM tooling. • Partner with Legal, and key stakeholder groups to translate business needs into clear technical requirements, solution designs, delivery plans, and measurable outcomes. • Own end-to-end delivery: discovery, architecture/design, build/configuration, testing/UAT, release, and production support-aligned with SDLC, change management, and operational readiness. • Establish and maintain integration patterns and data flows across core platforms (APIs, batch, middleware), supporting reporting, auditability, and downstream finance/procurement or records-management dependencies. • Drive strong engineering discipline: configuration standards, code quality, peer review, automated testing where applicable, environment management, and release governance. • Ensure embedded controls across workflows: approvals, access management, segregation of duties, evidence capture, and traceability aligned to audit and regulatory expectations. • Lead a small team of technical resources, providing coaching, prioritization, and delivery oversight; foster a culture of accountability and continuous improvement. • Manage vendors/partners as needed: solution selection support, implementation oversight, SLA/service performance, roadmap alignment, and cost management. • Provide transparent, executive-ready reporting on platform health, delivery status, risks/issues, dependencies, and remediation plans QUALIFICATIONS: 5+ years of progressive technology experience, including significant hands-on delivery experience supporting Legal systems and governance platforms in a complex environment. Experience in complex, regulated financial services; alternative asset management experience strongly preferred. Bachelor's degree required; Computer Science/Engineering background preferred. Strong written and verbal communication skills; proven ability to influence and collaborate across senior business and technology stakeholders. Demonstrated people leadership experience (direct management and/or leading teams through delivery and operations). Experience Required Strong engineering fundamentals and hands-on delivery experience, including: SQL (required), data analysis/troubleshooting, and comfort navigating complex process and data flows Experience with at least one general-purpose language (e.g., Python, Java, C#, JavaScript) and common integration approaches (REST APIs, services, batch) Working knowledge of SDLC, testing strategies, release/change management, and production support Experience implementing and supporting Legal and related risk platforms, including one or more of: Matter management / e-billing ecosystems (e.g., SimpleLegal, Passport, Apperio or similar) Legal Entity Management (e.g., GEMS or similar) CLM platforms (e.g., Icertis, Ironclad or similar) Intake and/or TPRM tooling (e.g., Zip and comparable tools) Experience with integrations and data/reporting enablement, including monitoring, data quality, and reconciliation practices. Demonstrated ability to deliver secure, audit-ready solutions with robust traceability, workflow transparency, and evidence retention. Proven ability to lead small teams and deliver results through a mix of direct technical contribution and team oversight. Vendor/SI management experience is a plus. Preferred Qualifications Familiarity applying automation and/or AI-enabled capabilities to legal operations (e.g., document classification, summarization, clause extraction, intake triage), with a strong governance and controls mindset. Experience establishing platform standards and scalable operating models (configuration governance, release discipline, support/runbooks). Strong process mapping and documentation skills (e.g., Lucidchart) to support requirements clarity and stakeholder alignment. General Requirements Ability to thrive in a fast-paced environment, manage competing priorities, and deliver high-quality outcomes. Strong ownership mindset with disciplined execution and attention to detail. Comfortable operating in ambiguity and translating evolving business needs into durable technical solutions. Collaborative, service-oriented approach; able to build credibility with Legal stakeholders and partner effectively across Corporate Technology. Willingness to flex hours as needed to support global stakeholders and production needs. Reporting Relationships Compensation The anticipated base salary range for this position is listed below. Total compensation may also include a discretionary performance-based bonus. Note, the range takes into account a broad spectrum of qualifications, including, but not limited to, years of relevant work experience, education, and other relevant qualifications specific to the role. $180,000 - $220,000 The firm also offers robust Benefits offerings. Ares U.S. Core Benefits include Comprehensive Medical/Rx, Dental and Vision plans; 401(k) program with company match; Flexible Savings Accounts (FSA); Healthcare Savings Accounts (HSA) with company contribution; Basic and Voluntary Life Insurance; Long-Term Disability (LTD) and Short-Term Disability (STD) insurance; Employee Assistance Program (EAP), and Commuter Benefits plan for parking and transit. Ares offers a number of additional benefits including access to a world-class medical advisory team, a mental health app that includes coaching, therapy and psychiatry, a mindfulness and wellbeing app, financial wellness benefit that includes access to a financial advisor, new parent leave, reproductive and adoption assistance, emergency backup care, matching gift program, education sponsorship program, and much more. There is no set deadline to apply for this job opportunity. Applications will be accepted on an ongoing basis until the search is no longer active.
09/23/2026
Full time
Over the last 20 years, Ares' success has been driven by our people and our culture. Today, our team is guided by our core values - Collaborative, Responsible, Entrepreneurial, Self-Aware, Trustworthy - and our purpose to be a catalyst for shared prosperity and a better future. Through our recruitment, career development and employee-focused programming, we are committed to fostering a welcoming and inclusive work environment where high-performance talent of diverse backgrounds, experiences, and perspectives can build careers within this exciting and growing industry. Job Description Ares is seeking a Legal Engineering Lead (Associate Vice President) to join Corporate Technology, reporting to the Head of Legal and Compliance Technology. This role will lead engineering delivery, platform health, and ongoing evolution of technology solutions supporting Ares' Legal and related risk and third-party governance workflows-driving scalable, secure, and audit-ready capabilities that improve operational efficiency and control effectiveness. The ideal candidate is a hands-on engineering leader with deep familiarity across Legal/GC processes and related risk workflows, ideally from a large alternative asset manager or similarly complex regulated firm (large law firm experience supporting sophisticated legal operations is also valued). This leader will bring strong technical fundamentals (e.g., SQL and at least one coding language), experience with integrations and data platforms, and the ability to partner closely with business stakeholders and vendors to deliver measurable outcomes. This role requires experience implementing and operating modern legal and risk platforms, including matter management and e-billing (e.g., Simple Legal / Passport / Apperio or similar), Legal Entity Management (e.g., GEMS), Contract Lifecycle Management (CLM) solutions (e.g., Icertis, Ironclad), and third-party risk management (TPRM) and intake tooling (e.g., Zip and comparable platforms). RESPONSIBILITIES: • Lead engineering delivery and operational ownership for the Legal and technology portfolio, ensuring solutions are scalable, resilient, secure, and audit-ready. • Serve as the technical lead for platform strategy, implementation, and support across Legal systems, including matter management/e-billing, legal entity management, CLM, and third-party intake/TPRM tooling. • Partner with Legal, and key stakeholder groups to translate business needs into clear technical requirements, solution designs, delivery plans, and measurable outcomes. • Own end-to-end delivery: discovery, architecture/design, build/configuration, testing/UAT, release, and production support-aligned with SDLC, change management, and operational readiness. • Establish and maintain integration patterns and data flows across core platforms (APIs, batch, middleware), supporting reporting, auditability, and downstream finance/procurement or records-management dependencies. • Drive strong engineering discipline: configuration standards, code quality, peer review, automated testing where applicable, environment management, and release governance. • Ensure embedded controls across workflows: approvals, access management, segregation of duties, evidence capture, and traceability aligned to audit and regulatory expectations. • Lead a small team of technical resources, providing coaching, prioritization, and delivery oversight; foster a culture of accountability and continuous improvement. • Manage vendors/partners as needed: solution selection support, implementation oversight, SLA/service performance, roadmap alignment, and cost management. • Provide transparent, executive-ready reporting on platform health, delivery status, risks/issues, dependencies, and remediation plans QUALIFICATIONS: 5+ years of progressive technology experience, including significant hands-on delivery experience supporting Legal systems and governance platforms in a complex environment. Experience in complex, regulated financial services; alternative asset management experience strongly preferred. Bachelor's degree required; Computer Science/Engineering background preferred. Strong written and verbal communication skills; proven ability to influence and collaborate across senior business and technology stakeholders. Demonstrated people leadership experience (direct management and/or leading teams through delivery and operations). Experience Required Strong engineering fundamentals and hands-on delivery experience, including: SQL (required), data analysis/troubleshooting, and comfort navigating complex process and data flows Experience with at least one general-purpose language (e.g., Python, Java, C#, JavaScript) and common integration approaches (REST APIs, services, batch) Working knowledge of SDLC, testing strategies, release/change management, and production support Experience implementing and supporting Legal and related risk platforms, including one or more of: Matter management / e-billing ecosystems (e.g., SimpleLegal, Passport, Apperio or similar) Legal Entity Management (e.g., GEMS or similar) CLM platforms (e.g., Icertis, Ironclad or similar) Intake and/or TPRM tooling (e.g., Zip and comparable tools) Experience with integrations and data/reporting enablement, including monitoring, data quality, and reconciliation practices. Demonstrated ability to deliver secure, audit-ready solutions with robust traceability, workflow transparency, and evidence retention. Proven ability to lead small teams and deliver results through a mix of direct technical contribution and team oversight. Vendor/SI management experience is a plus. Preferred Qualifications Familiarity applying automation and/or AI-enabled capabilities to legal operations (e.g., document classification, summarization, clause extraction, intake triage), with a strong governance and controls mindset. Experience establishing platform standards and scalable operating models (configuration governance, release discipline, support/runbooks). Strong process mapping and documentation skills (e.g., Lucidchart) to support requirements clarity and stakeholder alignment. General Requirements Ability to thrive in a fast-paced environment, manage competing priorities, and deliver high-quality outcomes. Strong ownership mindset with disciplined execution and attention to detail. Comfortable operating in ambiguity and translating evolving business needs into durable technical solutions. Collaborative, service-oriented approach; able to build credibility with Legal stakeholders and partner effectively across Corporate Technology. Willingness to flex hours as needed to support global stakeholders and production needs. Reporting Relationships Compensation The anticipated base salary range for this position is listed below. Total compensation may also include a discretionary performance-based bonus. Note, the range takes into account a broad spectrum of qualifications, including, but not limited to, years of relevant work experience, education, and other relevant qualifications specific to the role. $180,000 - $220,000 The firm also offers robust Benefits offerings. Ares U.S. Core Benefits include Comprehensive Medical/Rx, Dental and Vision plans; 401(k) program with company match; Flexible Savings Accounts (FSA); Healthcare Savings Accounts (HSA) with company contribution; Basic and Voluntary Life Insurance; Long-Term Disability (LTD) and Short-Term Disability (STD) insurance; Employee Assistance Program (EAP), and Commuter Benefits plan for parking and transit. Ares offers a number of additional benefits including access to a world-class medical advisory team, a mental health app that includes coaching, therapy and psychiatry, a mindfulness and wellbeing app, financial wellness benefit that includes access to a financial advisor, new parent leave, reproductive and adoption assistance, emergency backup care, matching gift program, education sponsorship program, and much more. There is no set deadline to apply for this job opportunity. Applications will be accepted on an ongoing basis until the search is no longer active.
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environment of a true industry disruptor, you'll find your place here. We value winning together-while learning, having fun, and making a profound difference for the dreamers and builders in the world. At DigitalOcean, Data Center Engineers play a critical role in building and operating the physical infrastructure that powers our cloud platform . Our team is responsible for deploying, maintaining, and scaling the servers and networking equipment that enable millions of developers to run their applications. From replacing a faulty drive to helping deploy infrastructure in a brand-new data center, our engineers work across the full lifecycle of hardware operations. You'll join a collaborative, fast-growing team with opportunities to work on large-scale deployments, new data center expansions, and next-generation infrastructure. This is a remote position. What You'll Be Doing Deploying, maintaining, and scaling DigitalOcean's data center infrastructure Racking, stacking, and cabling servers, power distribution units (PDUs), and network switches Installing and commissioning servers, networking equipment, storage systems, and GPU infrastructure Designing and executing large-scale infrastructure deployments including rack buildouts and cluster expansions Supporting new data center builds and expansions through rack layout planning and infrastructure readiness validation Deploying and troubleshooting high-density compute platforms including GPU and liquid-cooled infrastructure Diagnosing and repairing hardware issues across servers, networking equipment, and storage systems Debugging hardware, network, and Linux OS related issues Submitting RMAs and managing hardware lifecycle replacements with equipment vendors Managing shipping, receiving, and inventory tracking of data center hardware and components Maintaining deployment documentation, runbooks, and operational procedures Partnering with internal engineering teams to support hardware rollouts and infrastructure upgrades Supporting operational readiness for new data center deployments Participating in an on-call rotation to support infrastructure availability Mentoring junior engineers and contributing to team development What You'll Add to DigitalOcean 6-8+ years of experience deploying, operating, or supporting infrastructure within large-scale data center or cloud environments. Strong understanding of data center infrastructure including compute, networking, and storage systems Proficient understanding of data center power and cooling architecture Experience deploying and troubleshooting GPU hardware and high-density compute infrastructure Experience working with or supporting liquid-cooled infrastructure Strong Linux knowledge and command line experience; familiarity with bash scripting Strong troubleshooting skills across hardware, networking, and operating systems Understanding of data center cabling standards and deployment best practices Experience supporting large-scale infrastructure deployments or data center expansions Ability to analyze infrastructure incidents and drive operational improvements Strong problem-solving skills and the ability to perform deep technical investigations Experience maintaining deployment documentation and operational procedures Ability to collaborate with cross-functional engineering teams Willingness to participate in an on-call rotation Valid driver's license and ability to travel domestically and internationally (20-25%) Ability to lift objects up to 50 lbs and be on your feet for extended periods Ability to mentor and train junior engineers Compensation Range: $128,000 - $161,000 This is a hybrid role JR: Why You'll Like Working for DigitalOcean We innovate with purpose. You'll be a part of a cutting-edge technology company with an upward trajectory, who are proud to simplify cloud and AI so builders can spend more time creating software that changes the world. As a member of the team, you will be a Shark who thinks big, bold, and scrappy, like an owner with a bias for action and a powerful sense of responsibility for customers, products, employees, and decisions. We prioritize career development. At DO, you'll do the best work of your career. You will work with some of the smartest and most interesting people in the industry. We are a high-performance organization that will always challenge you to think big. Our organizational development team will provide you with resources to ensure you keep growing. We provide employees with reimbursement for relevant conferences, training, and education. All employees have access to LinkedIn Learning's 10,000+ courses to support their continued growth and development. We care about your well-being. Regardless of your location, we will provide you with a competitive array of benefits to support you from our Employee Assistance Program to Local Employee Meetups to flexible time off policy, to name a few. While the philosophy around our benefits is the same worldwide, specific benefits may vary based on local regulations and preferences. We reward our employees. The salary range for this position is based on market data, relevant years of experience, and skills. You may qualify for a bonus in addition to base salary; bonus amounts are determined based on company and individual performance. We also provide equity compensation to eligible employees, including equity grants upon hire and the option to participate in our Employee Stock Purchase Program. DigitalOcean is an equal-opportunity employer. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service. Application Limit: You may apply to a maximum of 3 positions within any 180-day period. This policy promotes better role-candidate matching and encourages thoughtful applications where your qualifications align most strongly.
09/23/2026
Full time
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environment of a true industry disruptor, you'll find your place here. We value winning together-while learning, having fun, and making a profound difference for the dreamers and builders in the world. At DigitalOcean, Data Center Engineers play a critical role in building and operating the physical infrastructure that powers our cloud platform . Our team is responsible for deploying, maintaining, and scaling the servers and networking equipment that enable millions of developers to run their applications. From replacing a faulty drive to helping deploy infrastructure in a brand-new data center, our engineers work across the full lifecycle of hardware operations. You'll join a collaborative, fast-growing team with opportunities to work on large-scale deployments, new data center expansions, and next-generation infrastructure. This is a remote position. What You'll Be Doing Deploying, maintaining, and scaling DigitalOcean's data center infrastructure Racking, stacking, and cabling servers, power distribution units (PDUs), and network switches Installing and commissioning servers, networking equipment, storage systems, and GPU infrastructure Designing and executing large-scale infrastructure deployments including rack buildouts and cluster expansions Supporting new data center builds and expansions through rack layout planning and infrastructure readiness validation Deploying and troubleshooting high-density compute platforms including GPU and liquid-cooled infrastructure Diagnosing and repairing hardware issues across servers, networking equipment, and storage systems Debugging hardware, network, and Linux OS related issues Submitting RMAs and managing hardware lifecycle replacements with equipment vendors Managing shipping, receiving, and inventory tracking of data center hardware and components Maintaining deployment documentation, runbooks, and operational procedures Partnering with internal engineering teams to support hardware rollouts and infrastructure upgrades Supporting operational readiness for new data center deployments Participating in an on-call rotation to support infrastructure availability Mentoring junior engineers and contributing to team development What You'll Add to DigitalOcean 6-8+ years of experience deploying, operating, or supporting infrastructure within large-scale data center or cloud environments. Strong understanding of data center infrastructure including compute, networking, and storage systems Proficient understanding of data center power and cooling architecture Experience deploying and troubleshooting GPU hardware and high-density compute infrastructure Experience working with or supporting liquid-cooled infrastructure Strong Linux knowledge and command line experience; familiarity with bash scripting Strong troubleshooting skills across hardware, networking, and operating systems Understanding of data center cabling standards and deployment best practices Experience supporting large-scale infrastructure deployments or data center expansions Ability to analyze infrastructure incidents and drive operational improvements Strong problem-solving skills and the ability to perform deep technical investigations Experience maintaining deployment documentation and operational procedures Ability to collaborate with cross-functional engineering teams Willingness to participate in an on-call rotation Valid driver's license and ability to travel domestically and internationally (20-25%) Ability to lift objects up to 50 lbs and be on your feet for extended periods Ability to mentor and train junior engineers Compensation Range: $128,000 - $161,000 This is a hybrid role JR: Why You'll Like Working for DigitalOcean We innovate with purpose. You'll be a part of a cutting-edge technology company with an upward trajectory, who are proud to simplify cloud and AI so builders can spend more time creating software that changes the world. As a member of the team, you will be a Shark who thinks big, bold, and scrappy, like an owner with a bias for action and a powerful sense of responsibility for customers, products, employees, and decisions. We prioritize career development. At DO, you'll do the best work of your career. You will work with some of the smartest and most interesting people in the industry. We are a high-performance organization that will always challenge you to think big. Our organizational development team will provide you with resources to ensure you keep growing. We provide employees with reimbursement for relevant conferences, training, and education. All employees have access to LinkedIn Learning's 10,000+ courses to support their continued growth and development. We care about your well-being. Regardless of your location, we will provide you with a competitive array of benefits to support you from our Employee Assistance Program to Local Employee Meetups to flexible time off policy, to name a few. While the philosophy around our benefits is the same worldwide, specific benefits may vary based on local regulations and preferences. We reward our employees. The salary range for this position is based on market data, relevant years of experience, and skills. You may qualify for a bonus in addition to base salary; bonus amounts are determined based on company and individual performance. We also provide equity compensation to eligible employees, including equity grants upon hire and the option to participate in our Employee Stock Purchase Program. DigitalOcean is an equal-opportunity employer. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service. Application Limit: You may apply to a maximum of 3 positions within any 180-day period. This policy promotes better role-candidate matching and encourages thoughtful applications where your qualifications align most strongly.
Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we're changing the way people think about and interact with personal finance. We're a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we're at the forefront. We're proud to come to work every day knowing that what we do has a direct impact on people's lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The Role You will be the technical leader for Digital Identity at SoFi under the Sofi Technology Solution Group: the platform group that powers identity, authorization, and entitlements for every product and every member across the company. Digital Identity runs Tier-0 infrastructure: the highest criticality rating at SoFi. Every product line, banking, lending, investing, credit cards, crypto depends on these platforms to know who a member is, what they're entitled to, and what they're authorized to do. When these platforms are down, SoFi is down. You'll define the technical strategy for this group. You'll architect solutions for complex, ambiguous problems: multi-person access patterns, cross-organizational platform convergence, and data integrity at financial-services scale. You'll build the engineering processes and culture that let a lean team operate Tier-0 infrastructure with confidence. And you'll push the boundaries of how we build, leveraging AI to accelerate development, prototype faster, and experiment with approaches that would have been impractical two years ago. What You'll Own Platform Technical Strategy Digital Identity operates multiple Tier-0 platforms spanning identity resolution, entitlement management, and fine-grained authorization. You own the technical strategy across all of them: setting the architectural direction, executing and leading designs, and ensuring the platforms evolve as a coherent system rather than independent services. Complex Authorization Architecture SoFi is expanding into scenarios where multiple people interact with shared financial resources: across business, family, and custodial contexts. You'll design the unified platform architecture that handles these patterns at scale: consistent access models, compliance-grade audit trails, and enforcement of regulatory requirements. This is one platform problem with many product surfaces. Cross-Organization Platform Convergence SoFi operates and integrates with multiple technology organizations with overlapping identity and authorization infrastructure. You'll lead the architectural vision for convergence: a shared platform primitives that multiple organizations consume while preserving the flexibility each product line needs. This requires navigating competing priorities, different technical stacks, and organizational boundaries. Operational Excellence & Data Integrity Tier-0 financial platforms demand more than uptime. You'll architect the verification and reconciliation systems that prove these platforms are correct: automated integrity checks, drift detection, and self-healing mechanisms. You'll establish the operational processes, incident response standards, and reliability practices that let the team ship with confidence and sleep at night. Engineering Culture & Team Uplift You'll raise the bar for how this team builds software. That means establishing rigorous design review processes, defining engineering standards that compound over time, mentoring senior ICs into technical leaders, and creating the feedback loops that turn incidents into prevention. You're not just the best engineer on the team: you're the reason the whole team gets better. Strategic Investment Identification You won't just execute on the roadmap handed to you. You'll identify the next set of high-leverage technical investments: where the platforms should go, what capabilities are missing, which emerging patterns (in authorization, in AI, in infrastructure) should be adopted before the business asks for them. What We're Looking For Required: Distributed systems architecture at scale. You've designed and shipped platforms that other engineering teams depend on: not just consumed services, but built them. You understand the failure modes of event-driven systems, eventual consistency, and cross-service data integrity. You've made hard tradeoffs between consistency, availability, and latency in production. Technical leadership with accountability built in. You don't just design systems: you design systems that prove they're correct. Reconciliation mechanisms, audit trails, integrity guarantees, automated verification. You've built infrastructure where "trust but verify" is architecture, not process. AI fluency and innovation. You actively use AI to build, prototype, and experiment. You've integrated AI-assisted development into your workflow and can articulate where it accelerates engineering and where it introduces risk. You push teams to adopt AI-native approaches to development not as a novelty, but as a competitive advantage in velocity and experimentation. Group-level influence and execution. You've driven technical strategy across multiple teams. You've navigated ambiguity where business goals were clear but the right technical problems to solve were not. You've represented your organization's technical direction to peer groups and senior / executive leadership. Engineering culture builder. You've established processes, standards, and practices that made entire teams more effective and not just shipped features yourself. You care about design review rigor, operational readiness, on-call excellence, and mentoring senior engineers into technical leaders. Ownership of outcomes, not just systems. You measure your work by what it enabled such as products shipped, risks eliminated, teams unblocked and not by the complexity of what you built. Preferred: Experience building identity & authorization platforms especially in multi-tenant or consumer-facing contexts. Familiarity with relationship-based access control models, fine-grained authorization systems, or identity federation infrastructure. Experience in financial services or regulated industries where compliance, audit trails, and data integrity are architectural requirements, not afterthoughts. Track record of platform convergence: merging or unifying infrastructure across acquisitions, subsidiaries, or organizations with different technical stacks. What This Isn't: This is not a role for someone who wants to write code in a corner. You'll spend significant time on architecture, cross-team alignment, design reviews, and mentorship. You will write code and drive credibility with your depth but the highest-leverage output is the technical direction you set, the engineering culture you build, and the engineers you develop. Why This Role? Scope. Digital Identity is horizontal infrastructure. Every product at SoFi depends on your platforms. A design decision you make affects millions of members. Hard problems. Multi-person financial authorization, cross-org platform convergence, data integrity at Tier-0 scale. These aren't optimizations: they're greenfield architecture for a company-wide platform. Visibility. Tier-0 means executives know when you ship and when you don't. Impact is not abstract here. Build + Lead. Deep distributed systems work and Group-level technical leadership. You'll build both muscles, every week. You'll architect systems and build the team's engineering culture in equal measure. AI-forward engineering. We're not waiting for the industry to figure out how AI changes platform engineering. You'll help define that for Digital Identity, using AI to move faster, experiment more, and build systems that would have been impractical with traditional approaches alone. Compensation and Benefits The base pay range for this role is listed below. Final base pay offer will be determined based on individual factors such as the candidate's experience, skills, and location. To view all of our comprehensive and competitive benefits, visit our Benefits at SoFi page! SoFi provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion (including religious dress and grooming practices), sex (including pregnancy, childbirth and related medical conditions, breastfeeding, and conditions related to breastfeeding), gender, gender identity, gender expression, national origin, ancestry, age (40 or over), physical or medical disability, medical condition, marital status, registered domestic partner status, sexual orientation, genetic information, military and/or veteran status, or any other basis prohibited by applicable state or federal law. The Company hires the best qualified candidate for the job, without regard to protected characteristics. Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. New York applicants: Notice of Employee Rights SoFi is committed to an inclusive culture. As part of this commitment, SoFi offers reasonable accommodations to candidates with physical or mental disabilities . click apply for full job details
09/23/2026
Full time
Employee Applicant Privacy Notice Who we are: Shape a brighter financial future with us. Together with our members, we're changing the way people think about and interact with personal finance. We're a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we're at the forefront. We're proud to come to work every day knowing that what we do has a direct impact on people's lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world. The Role You will be the technical leader for Digital Identity at SoFi under the Sofi Technology Solution Group: the platform group that powers identity, authorization, and entitlements for every product and every member across the company. Digital Identity runs Tier-0 infrastructure: the highest criticality rating at SoFi. Every product line, banking, lending, investing, credit cards, crypto depends on these platforms to know who a member is, what they're entitled to, and what they're authorized to do. When these platforms are down, SoFi is down. You'll define the technical strategy for this group. You'll architect solutions for complex, ambiguous problems: multi-person access patterns, cross-organizational platform convergence, and data integrity at financial-services scale. You'll build the engineering processes and culture that let a lean team operate Tier-0 infrastructure with confidence. And you'll push the boundaries of how we build, leveraging AI to accelerate development, prototype faster, and experiment with approaches that would have been impractical two years ago. What You'll Own Platform Technical Strategy Digital Identity operates multiple Tier-0 platforms spanning identity resolution, entitlement management, and fine-grained authorization. You own the technical strategy across all of them: setting the architectural direction, executing and leading designs, and ensuring the platforms evolve as a coherent system rather than independent services. Complex Authorization Architecture SoFi is expanding into scenarios where multiple people interact with shared financial resources: across business, family, and custodial contexts. You'll design the unified platform architecture that handles these patterns at scale: consistent access models, compliance-grade audit trails, and enforcement of regulatory requirements. This is one platform problem with many product surfaces. Cross-Organization Platform Convergence SoFi operates and integrates with multiple technology organizations with overlapping identity and authorization infrastructure. You'll lead the architectural vision for convergence: a shared platform primitives that multiple organizations consume while preserving the flexibility each product line needs. This requires navigating competing priorities, different technical stacks, and organizational boundaries. Operational Excellence & Data Integrity Tier-0 financial platforms demand more than uptime. You'll architect the verification and reconciliation systems that prove these platforms are correct: automated integrity checks, drift detection, and self-healing mechanisms. You'll establish the operational processes, incident response standards, and reliability practices that let the team ship with confidence and sleep at night. Engineering Culture & Team Uplift You'll raise the bar for how this team builds software. That means establishing rigorous design review processes, defining engineering standards that compound over time, mentoring senior ICs into technical leaders, and creating the feedback loops that turn incidents into prevention. You're not just the best engineer on the team: you're the reason the whole team gets better. Strategic Investment Identification You won't just execute on the roadmap handed to you. You'll identify the next set of high-leverage technical investments: where the platforms should go, what capabilities are missing, which emerging patterns (in authorization, in AI, in infrastructure) should be adopted before the business asks for them. What We're Looking For Required: Distributed systems architecture at scale. You've designed and shipped platforms that other engineering teams depend on: not just consumed services, but built them. You understand the failure modes of event-driven systems, eventual consistency, and cross-service data integrity. You've made hard tradeoffs between consistency, availability, and latency in production. Technical leadership with accountability built in. You don't just design systems: you design systems that prove they're correct. Reconciliation mechanisms, audit trails, integrity guarantees, automated verification. You've built infrastructure where "trust but verify" is architecture, not process. AI fluency and innovation. You actively use AI to build, prototype, and experiment. You've integrated AI-assisted development into your workflow and can articulate where it accelerates engineering and where it introduces risk. You push teams to adopt AI-native approaches to development not as a novelty, but as a competitive advantage in velocity and experimentation. Group-level influence and execution. You've driven technical strategy across multiple teams. You've navigated ambiguity where business goals were clear but the right technical problems to solve were not. You've represented your organization's technical direction to peer groups and senior / executive leadership. Engineering culture builder. You've established processes, standards, and practices that made entire teams more effective and not just shipped features yourself. You care about design review rigor, operational readiness, on-call excellence, and mentoring senior engineers into technical leaders. Ownership of outcomes, not just systems. You measure your work by what it enabled such as products shipped, risks eliminated, teams unblocked and not by the complexity of what you built. Preferred: Experience building identity & authorization platforms especially in multi-tenant or consumer-facing contexts. Familiarity with relationship-based access control models, fine-grained authorization systems, or identity federation infrastructure. Experience in financial services or regulated industries where compliance, audit trails, and data integrity are architectural requirements, not afterthoughts. Track record of platform convergence: merging or unifying infrastructure across acquisitions, subsidiaries, or organizations with different technical stacks. What This Isn't: This is not a role for someone who wants to write code in a corner. You'll spend significant time on architecture, cross-team alignment, design reviews, and mentorship. You will write code and drive credibility with your depth but the highest-leverage output is the technical direction you set, the engineering culture you build, and the engineers you develop. Why This Role? Scope. Digital Identity is horizontal infrastructure. Every product at SoFi depends on your platforms. A design decision you make affects millions of members. Hard problems. Multi-person financial authorization, cross-org platform convergence, data integrity at Tier-0 scale. These aren't optimizations: they're greenfield architecture for a company-wide platform. Visibility. Tier-0 means executives know when you ship and when you don't. Impact is not abstract here. Build + Lead. Deep distributed systems work and Group-level technical leadership. You'll build both muscles, every week. You'll architect systems and build the team's engineering culture in equal measure. AI-forward engineering. We're not waiting for the industry to figure out how AI changes platform engineering. You'll help define that for Digital Identity, using AI to move faster, experiment more, and build systems that would have been impractical with traditional approaches alone. Compensation and Benefits The base pay range for this role is listed below. Final base pay offer will be determined based on individual factors such as the candidate's experience, skills, and location. To view all of our comprehensive and competitive benefits, visit our Benefits at SoFi page! SoFi provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion (including religious dress and grooming practices), sex (including pregnancy, childbirth and related medical conditions, breastfeeding, and conditions related to breastfeeding), gender, gender identity, gender expression, national origin, ancestry, age (40 or over), physical or medical disability, medical condition, marital status, registered domestic partner status, sexual orientation, genetic information, military and/or veteran status, or any other basis prohibited by applicable state or federal law. The Company hires the best qualified candidate for the job, without regard to protected characteristics. Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. New York applicants: Notice of Employee Rights SoFi is committed to an inclusive culture. As part of this commitment, SoFi offers reasonable accommodations to candidates with physical or mental disabilities . click apply for full job details