it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

71 jobs found

Email me jobs like this
Refine Search
Current Search
manager digital transformation
Source to Contract Associate - Procurement Services (SaaS/Tech)
WNS Denali, Powered by the Smart Cube Pittsburgh, Pennsylvania
Job DescriptionJob DescriptionCompany Description WNS Procurement, part of Capgemini, is an Agentic AI-powered leader in intelligent operations and transformation, serving more than 700 clients across 10 industries, including Banking and Financial Services, Healthcare, Insurance, Shipping and Logistics, and Travel and Hospitality. We combine our deep industry knowledge with technology and analytics expertise to co-create innovative, digital-led transformational solutions with clients across multiple industries. WNS Procurement, a strategic business unit within WNS, Part of Capgemini is a market leader in procurement transformation & advisory, managed services, intelligence and analytics, and digital tools. Our mission is to enable procurement to become the top value creator in the business by implementing transformational operating models that are category-driven, insight-led, and digitally enabled. Why Join WNS Procurement? Client-Centric Approach: Help clients achieve their business goals by implementing customized, next-generation procurement solutions. Collaborative Culture: Join a diverse and inclusive workplace where teamwork and collaboration are at the heart of everything we do. Innovative Environment: Be part of a team that leverages cutting-edge technology and data-driven insights to revolutionize procurement processes. Global Impact: Work with leading global companies and make a significant impact on their procurement strategies. Career Growth: We offer extensive professional development opportunities, ensuring that you grow alongside the company. Job Description ob Title: Source to Contract Associate Location: Remote Employment Type: Fill Time Industry: Procurement Services - SaaS/Tech Experience Level: Entry-Level About the Role We are seeking a Source to Contract Associate to join our team and contribute to a client sourcing service delivery team, managing the development and execution of sourcing projects of various complexity levels. The Sourcing Associate is responsible for executing simple to complex sourcing projects, from start to finish, including but not limited to - validating sourcing requests per client guidelines, analyzing requirements for the bid packages, communicating with client requesters to close any data gaps, preparing and publishing the package to the identified suppliers for bidding purposes, following up with suppliers to ensure adequate participation, analyzing and summarizing proposals, coordinating evaluation of suppliers and facilitating decision makings. The ideal candidate will possess Source to Contract experience and skills: negotiation of software contracts, process project management, effective communication skills. This person should also be self driven, independent, resourceful and adaptive. Key Responsibilities Sourcing/Contract Execution: Provides sourcing services (RFx, reverse auctions, negotiations) to our clients based on predefined service levels, managing 5-12 simple to complex projects simultaneously with limited supervision/guidance or independently Once a project request is received, assess completeness of requirements, perform due diligence and validation, follow up with the requestor to obtain further project details, answer any process related questions, and provide updates on project status Lead and execute the WNS sourcing process from start to finish, e.g., spend data collection and analysis, supply market analysis and supplier discovery if needed, RFx development, publishing and management, reverse auctions, negotiations, bid response analysis and synthesis Complete negotiations in software environment and credit based system enterprise licensing. Tech, SaaS, Professional Services, and Corporate Services deals. Manage suppliers during the sourcing process, including conducting supplier orientation and training on the sourcing process, following up with suppliers on milestone activities Perform in depth analysis of proposals submitted, facilitate supplier evaluations, draw conclusions, prepare comprehensive summaries, and present to the client in a concise manner Process Improvement Develop and/or customize various templates such as RFx, pricing, and proposal evaluation Identify areas of process inefficiencies and suggest improvements to management to improve efficiency and effectiveness of Denali processes Client and Project Management Develop and maintain comprehensive project documentation Collaborate with Client Category Managers and/or WNS SMEs on project strategy and determine sourcing approach and methodology to be applied based on the strategy Qualifications Required Qualifications Preferred Qualifications Bachelor's Degree Minimum 2-3 years of work experience in procurement, sourcing, operations, or supply chain Procurement Functional/Technical Skills: Strong Sourcing skills, including sourcing methodology and process, eSourcing technology, RFX development and management Familiarity with various best practice sourcing approaches and techniques Strong analytical skills and ability to work with data in Excel using pivot tables and various other key Excel functions to analyze, organize and present data Ability to synthesize data, draw conclusion, and make recommendations based on understanding of client objectives and requirements Proficiency of one or more eSourcing tools (e.g., Ariba, Procuri, Iasta, Zycus, ERP eSourcing module) CLAUDE IS PREFERRED Basic to intermediate negotiation skills in software environment and credit based system enterprise licensing. Tech, SaaS, Professional Services, and Corporate Services deals. Understanding of retail store specific categories and their pricing/cost structures, terminology, key vendors, key requirements, typical sourcing levers, etc. Business Fundamentals: Excellent written and verbal communication skills Demonstrated teamwork Excellent project management skills including project planning, time management, multitasking, critical path definition Client Services Capabilities: Strong customer service orientation including demonstrated issue resolution and relationship management skills Ability to learn and master client specific processes, terminology, political environment, systems and unique requirements by various business groups Solid decision making ability using available facts in sensitive client situations Additional information Compensation Disclosure The base salary range for this position is $75K - $90K annually. This range reflects the base pay that we reasonably expect to offer for the role across our hiring locations. Final compensation will be determined based on a combination of factors, including but not limited to: Geographic location (state and city of residence) Overall professional experience Directly relevant experience Education and certifications Industry knowledge and expertise Skills and competencies In addition to base pay, this role may be eligible for performance-based bonuses, or incentive pay, or commissions, which are not included in the listed base salary range. WNS complies with all applicable federal, state, and local pay transparency laws, including those in California, Colorado, New York, Washington, and Illinois. Where required by law, we will provide additional details about compensation and benefits to qualified applicants. Benefits Overview Our benefits package includes (but is not limited to): - Medical, dental, and vision insurance - Paid time off (PTO), holidays, and sick leave - 401(k) with company match or other retirement plan - Life and AD&D Insurance - Employee Assistance Program Equal Opportunity Employer Statement WNS is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity or expression, national origin, age, disability, genetic information, veteran status, or any other status protected under federal, state, or local law. We also provide reasonable accommodations to individuals with disabilities and for sincerely held religious beliefs in all aspects of employment, including the application process.
09/23/2026
Full time
Job DescriptionJob DescriptionCompany Description WNS Procurement, part of Capgemini, is an Agentic AI-powered leader in intelligent operations and transformation, serving more than 700 clients across 10 industries, including Banking and Financial Services, Healthcare, Insurance, Shipping and Logistics, and Travel and Hospitality. We combine our deep industry knowledge with technology and analytics expertise to co-create innovative, digital-led transformational solutions with clients across multiple industries. WNS Procurement, a strategic business unit within WNS, Part of Capgemini is a market leader in procurement transformation & advisory, managed services, intelligence and analytics, and digital tools. Our mission is to enable procurement to become the top value creator in the business by implementing transformational operating models that are category-driven, insight-led, and digitally enabled. Why Join WNS Procurement? Client-Centric Approach: Help clients achieve their business goals by implementing customized, next-generation procurement solutions. Collaborative Culture: Join a diverse and inclusive workplace where teamwork and collaboration are at the heart of everything we do. Innovative Environment: Be part of a team that leverages cutting-edge technology and data-driven insights to revolutionize procurement processes. Global Impact: Work with leading global companies and make a significant impact on their procurement strategies. Career Growth: We offer extensive professional development opportunities, ensuring that you grow alongside the company. Job Description ob Title: Source to Contract Associate Location: Remote Employment Type: Fill Time Industry: Procurement Services - SaaS/Tech Experience Level: Entry-Level About the Role We are seeking a Source to Contract Associate to join our team and contribute to a client sourcing service delivery team, managing the development and execution of sourcing projects of various complexity levels. The Sourcing Associate is responsible for executing simple to complex sourcing projects, from start to finish, including but not limited to - validating sourcing requests per client guidelines, analyzing requirements for the bid packages, communicating with client requesters to close any data gaps, preparing and publishing the package to the identified suppliers for bidding purposes, following up with suppliers to ensure adequate participation, analyzing and summarizing proposals, coordinating evaluation of suppliers and facilitating decision makings. The ideal candidate will possess Source to Contract experience and skills: negotiation of software contracts, process project management, effective communication skills. This person should also be self driven, independent, resourceful and adaptive. Key Responsibilities Sourcing/Contract Execution: Provides sourcing services (RFx, reverse auctions, negotiations) to our clients based on predefined service levels, managing 5-12 simple to complex projects simultaneously with limited supervision/guidance or independently Once a project request is received, assess completeness of requirements, perform due diligence and validation, follow up with the requestor to obtain further project details, answer any process related questions, and provide updates on project status Lead and execute the WNS sourcing process from start to finish, e.g., spend data collection and analysis, supply market analysis and supplier discovery if needed, RFx development, publishing and management, reverse auctions, negotiations, bid response analysis and synthesis Complete negotiations in software environment and credit based system enterprise licensing. Tech, SaaS, Professional Services, and Corporate Services deals. Manage suppliers during the sourcing process, including conducting supplier orientation and training on the sourcing process, following up with suppliers on milestone activities Perform in depth analysis of proposals submitted, facilitate supplier evaluations, draw conclusions, prepare comprehensive summaries, and present to the client in a concise manner Process Improvement Develop and/or customize various templates such as RFx, pricing, and proposal evaluation Identify areas of process inefficiencies and suggest improvements to management to improve efficiency and effectiveness of Denali processes Client and Project Management Develop and maintain comprehensive project documentation Collaborate with Client Category Managers and/or WNS SMEs on project strategy and determine sourcing approach and methodology to be applied based on the strategy Qualifications Required Qualifications Preferred Qualifications Bachelor's Degree Minimum 2-3 years of work experience in procurement, sourcing, operations, or supply chain Procurement Functional/Technical Skills: Strong Sourcing skills, including sourcing methodology and process, eSourcing technology, RFX development and management Familiarity with various best practice sourcing approaches and techniques Strong analytical skills and ability to work with data in Excel using pivot tables and various other key Excel functions to analyze, organize and present data Ability to synthesize data, draw conclusion, and make recommendations based on understanding of client objectives and requirements Proficiency of one or more eSourcing tools (e.g., Ariba, Procuri, Iasta, Zycus, ERP eSourcing module) CLAUDE IS PREFERRED Basic to intermediate negotiation skills in software environment and credit based system enterprise licensing. Tech, SaaS, Professional Services, and Corporate Services deals. Understanding of retail store specific categories and their pricing/cost structures, terminology, key vendors, key requirements, typical sourcing levers, etc. Business Fundamentals: Excellent written and verbal communication skills Demonstrated teamwork Excellent project management skills including project planning, time management, multitasking, critical path definition Client Services Capabilities: Strong customer service orientation including demonstrated issue resolution and relationship management skills Ability to learn and master client specific processes, terminology, political environment, systems and unique requirements by various business groups Solid decision making ability using available facts in sensitive client situations Additional information Compensation Disclosure The base salary range for this position is $75K - $90K annually. This range reflects the base pay that we reasonably expect to offer for the role across our hiring locations. Final compensation will be determined based on a combination of factors, including but not limited to: Geographic location (state and city of residence) Overall professional experience Directly relevant experience Education and certifications Industry knowledge and expertise Skills and competencies In addition to base pay, this role may be eligible for performance-based bonuses, or incentive pay, or commissions, which are not included in the listed base salary range. WNS complies with all applicable federal, state, and local pay transparency laws, including those in California, Colorado, New York, Washington, and Illinois. Where required by law, we will provide additional details about compensation and benefits to qualified applicants. Benefits Overview Our benefits package includes (but is not limited to): - Medical, dental, and vision insurance - Paid time off (PTO), holidays, and sick leave - 401(k) with company match or other retirement plan - Life and AD&D Insurance - Employee Assistance Program Equal Opportunity Employer Statement WNS is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity or expression, national origin, age, disability, genetic information, veteran status, or any other status protected under federal, state, or local law. We also provide reasonable accommodations to individuals with disabilities and for sincerely held religious beliefs in all aspects of employment, including the application process.
Branch Operations Manager, Lebanon, PA
Santander Holdings USA Inc Lebanon, Pennsylvania
It Starts Here: Santander is a global leader and innovator in the financial services industry and is evolving from a high-impact brand into a technology-driven organization. Our people are at the heart of this journey and together, we are driving a customer-centric transformation that values bold thinking, innovation, and the courage to challenge what's possible. This is more than a strategic shift. It's a chance for driven professionals to grow, learn, and make a real difference. If you are interested in exploring the possibilities We Want to Talk to You! The Difference You Make: As a Branch Operations Manager, you ensure the branch operates efficiently and securely while delivering exceptional customer experiences and fostering team member growth. You oversee risk controls by ensuring compliance with policies, procedures, and regulatory requirements, minimizing operational risks tied to cash handling and transactions. This role includes enhancing the customer experience by ensuring smooth transaction processing, resolving issues promptly, lobby management and creating a welcoming environment. You serve as a trusted expert, providing clarity on policies, guidance on execution and assistance with escalations. This position will provide support for the Lebanon Plaza branch and Hershey Cocoa branches. Assist customers with various transactions, including deposits, withdrawals and payments. Oversee operational risk control measures to safeguard branch assets, including Vault and ATM custodianship. Ensure an elevated customer experience, delivering personalized, seamless, and attentive service. Effective lobby management to optimize customer flow and engagement. Resolve customer issues promptly and effectively. Build and maintain strong relationships with customers to elevate their banking experience and foster loyalty. Engage customers through digital platforms to enhance customer interactions and educate them on self-service options. Conduct cash counts and maintain accurate audit logs. Support the teller line, use coaching tools, and provide feedback to ensure efficient and accurate transactions. Communicate clearly and effectively with customers in person, over the phone, or through digital channels. Utilize data-driven decision-making to improve branch performance and operational efficiency. Assist colleagues in achieving their developmental goals and career aspirations. Responsibilities may extend to supporting nearby branch locations based on business necessity or as required based on branch designation. What You Bring: To perform this job successfully, an individual must be able to perform each essential duty satisfactorily. The requirements listed below are representative of the knowledge, skill, and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. Education: High school diploma, GED: or equivalent work experience - Required Qualifications: 3+ Years Demonstrated successful experience in branch banking or a related operations/support function - Required. (OR) 12+ Months Demonstrated successful Santander experience related to the essential functions and responsibilities of the Branch Operations Manager role. District Executive, District Operations Manager and Region President endorsement of performance - Required. (AND) 18+ Months Cash handling experience - Required. (AND) 18+ Months Customer service experience within a high volume, fast paced and constantly changing environment. - Required. Proficient in cash handling and maintaining audit logs. Excellent customer service skills and a passion for helping others. Proven ability to build relationships and enhance customer experience. Strong problem-solving skills with a proactive approach to issue resolution. Proficient in using digital tools and technology to enhance customer engagement. Ability to make data-driven decisions to improve operational outcomes. Strong knowledge of company policy, compliance regulations, risk management and loss prevention. Ability to work in a fast-paced environment and manage multiple priorities. Excellent communication, consultative and influence skills both verbal and written. Self-motivated to succeed in a goal driven environment. Ability to interact with integrity and professionalism with customers and employees. Computer proficiency and basic math skills. Ability to work branch hours, which can include weekends and evenings. Certifications: No Certifications listed for this job. It Would Be Nice For You To Have: Established work history or equivalent demonstrated through a combination of work experience, training, military service, or education. Preferred experience in Microsoft Office products. Work Authorization & Sponsorship: Applicants must be legally authorized to work in the United States on a full-time basis without requiring employer sponsorship to commence employment. What Else You Need To Know: The base pay range for this position is posted below and represents the annualized salary range. For hourly positions (non-exempt), the annual range is based on a 40-hour work week. The exact compensation may vary based on skills, experience, training, licensure and certifications and location. Base Pay Range: Minimum: $38,250.00 USD Maximum: $64,000.00 USD We Value Your Impact: Your contribution matters and it's recognized. You can expect a fair and competitive rewards package that reflects the impact you create and the value you deliver. We know rewards go beyond numbers. Offering more than just a paycheck our benefits are designed to support you, your family and your well-being, now and into the future. Santander Benefits - 2026 Santander OnGoing/NH eGuide () Risk Culture: We embrace a strong risk culture and all of our professionals at all levels are expected to take a proactive and responsible approach toward risk management. EEO Statement: At Santander, we value and respect differences in our workforce. We actively encourage everyone to apply. Santander is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, genetics, disability, age, veteran status or any other characteristic protected by law. Working Conditions: Frequent minimal physical effort such as sitting, standing and walking is required for this role. Depending on location, occasional moving and lifting light equipment and/or furniture may be required. Employer Rights: This job description does not list all of the job duties of the job. You may be asked by your supervisors or managers to perform other duties. You may be evaluated in part based upon your performance of the tasks listed in this job description. The employer has the right to revise this job description at any time. This job description is not a contract for employment and either you or the employer may terminate your employment at any time for any reason. What To Do Next : If this sounds like a role you are interested in, then please apply. We are committed to providing an inclusive and accessible application process for all candidates. If you require any assistance or accommodation due to a disability or any other reason, please contact us at to discuss your needs.
09/23/2026
Full time
It Starts Here: Santander is a global leader and innovator in the financial services industry and is evolving from a high-impact brand into a technology-driven organization. Our people are at the heart of this journey and together, we are driving a customer-centric transformation that values bold thinking, innovation, and the courage to challenge what's possible. This is more than a strategic shift. It's a chance for driven professionals to grow, learn, and make a real difference. If you are interested in exploring the possibilities We Want to Talk to You! The Difference You Make: As a Branch Operations Manager, you ensure the branch operates efficiently and securely while delivering exceptional customer experiences and fostering team member growth. You oversee risk controls by ensuring compliance with policies, procedures, and regulatory requirements, minimizing operational risks tied to cash handling and transactions. This role includes enhancing the customer experience by ensuring smooth transaction processing, resolving issues promptly, lobby management and creating a welcoming environment. You serve as a trusted expert, providing clarity on policies, guidance on execution and assistance with escalations. This position will provide support for the Lebanon Plaza branch and Hershey Cocoa branches. Assist customers with various transactions, including deposits, withdrawals and payments. Oversee operational risk control measures to safeguard branch assets, including Vault and ATM custodianship. Ensure an elevated customer experience, delivering personalized, seamless, and attentive service. Effective lobby management to optimize customer flow and engagement. Resolve customer issues promptly and effectively. Build and maintain strong relationships with customers to elevate their banking experience and foster loyalty. Engage customers through digital platforms to enhance customer interactions and educate them on self-service options. Conduct cash counts and maintain accurate audit logs. Support the teller line, use coaching tools, and provide feedback to ensure efficient and accurate transactions. Communicate clearly and effectively with customers in person, over the phone, or through digital channels. Utilize data-driven decision-making to improve branch performance and operational efficiency. Assist colleagues in achieving their developmental goals and career aspirations. Responsibilities may extend to supporting nearby branch locations based on business necessity or as required based on branch designation. What You Bring: To perform this job successfully, an individual must be able to perform each essential duty satisfactorily. The requirements listed below are representative of the knowledge, skill, and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. Education: High school diploma, GED: or equivalent work experience - Required Qualifications: 3+ Years Demonstrated successful experience in branch banking or a related operations/support function - Required. (OR) 12+ Months Demonstrated successful Santander experience related to the essential functions and responsibilities of the Branch Operations Manager role. District Executive, District Operations Manager and Region President endorsement of performance - Required. (AND) 18+ Months Cash handling experience - Required. (AND) 18+ Months Customer service experience within a high volume, fast paced and constantly changing environment. - Required. Proficient in cash handling and maintaining audit logs. Excellent customer service skills and a passion for helping others. Proven ability to build relationships and enhance customer experience. Strong problem-solving skills with a proactive approach to issue resolution. Proficient in using digital tools and technology to enhance customer engagement. Ability to make data-driven decisions to improve operational outcomes. Strong knowledge of company policy, compliance regulations, risk management and loss prevention. Ability to work in a fast-paced environment and manage multiple priorities. Excellent communication, consultative and influence skills both verbal and written. Self-motivated to succeed in a goal driven environment. Ability to interact with integrity and professionalism with customers and employees. Computer proficiency and basic math skills. Ability to work branch hours, which can include weekends and evenings. Certifications: No Certifications listed for this job. It Would Be Nice For You To Have: Established work history or equivalent demonstrated through a combination of work experience, training, military service, or education. Preferred experience in Microsoft Office products. Work Authorization & Sponsorship: Applicants must be legally authorized to work in the United States on a full-time basis without requiring employer sponsorship to commence employment. What Else You Need To Know: The base pay range for this position is posted below and represents the annualized salary range. For hourly positions (non-exempt), the annual range is based on a 40-hour work week. The exact compensation may vary based on skills, experience, training, licensure and certifications and location. Base Pay Range: Minimum: $38,250.00 USD Maximum: $64,000.00 USD We Value Your Impact: Your contribution matters and it's recognized. You can expect a fair and competitive rewards package that reflects the impact you create and the value you deliver. We know rewards go beyond numbers. Offering more than just a paycheck our benefits are designed to support you, your family and your well-being, now and into the future. Santander Benefits - 2026 Santander OnGoing/NH eGuide () Risk Culture: We embrace a strong risk culture and all of our professionals at all levels are expected to take a proactive and responsible approach toward risk management. EEO Statement: At Santander, we value and respect differences in our workforce. We actively encourage everyone to apply. Santander is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, genetics, disability, age, veteran status or any other characteristic protected by law. Working Conditions: Frequent minimal physical effort such as sitting, standing and walking is required for this role. Depending on location, occasional moving and lifting light equipment and/or furniture may be required. Employer Rights: This job description does not list all of the job duties of the job. You may be asked by your supervisors or managers to perform other duties. You may be evaluated in part based upon your performance of the tasks listed in this job description. The employer has the right to revise this job description at any time. This job description is not a contract for employment and either you or the employer may terminate your employment at any time for any reason. What To Do Next : If this sounds like a role you are interested in, then please apply. We are committed to providing an inclusive and accessible application process for all candidates. If you require any assistance or accommodation due to a disability or any other reason, please contact us at to discuss your needs.
Change Control & Configuration Manager - Intermediate (Fed Gov)
RE/SPEC Inc. Reston, Virginia
Job Description Job Description Company Description Big challenges need bold thinkers. If you're someone who sees problems as opportunities, you'll thrive here. RESPEC is 100% employee-owned, which means we take ownership of every challenge. Here, your ideas drive real solutions. Since 1969, we've tackled complex challenges in energy transition, infrastructure resilience, digital transformation, and sustainability. At RESPEC, you'll work alongside clients to take on critical problems. Depending on your expertise, you might design infrastructure in remote locations, develop renewable energy solutions for global projects, or apply data-driven technology to improve mining and water systems. We bring deep technical knowledge, real-world experience, and a commitment to work that matters. If you're looking for a place where your contributions have real impact, you'll fit right in. We do not accept unsolicited resumes from third-party recruiters. Job Description RESPEC is looking to hire an experienced Change Control & Configuration Manager - Intermediate to support Indian Affairs (IA) and Office of Information Technology (OIT) programs, projects, and operational initiatives. The Change Control & Configuration Manager provides specialized expertise in configuration management, change control, and release management while managing assigned IT projects throughout the project lifecycle. This position serves as an intermediate-level Project Manager and coordinates Government, business, technical, security, vendor, and contractor stakeholders to manage scope, schedule, cost, quality, risk, configuration, change, documentation, compliance, and delivery performance. The position is responsible for maintaining effective Change and Configuration Management (CCM) processes, ensuring configuration information and records remain accurate and current, coordinating change and release activities, supporting Configuration Control Board (CCB) activities, and helping ensure compliance with applicable NIST, Department of the Interior (DOI), Indian Affairs (IA), and OIT requirements. In addition to Change and Configuration Management expertise, the successful candidate must bring substantive experience in at least one complementary IT delivery discipline: Quality & Test Management, IT Asset Management, or Business Analysis & Requirements Management. Working with the Contracting Officer's Representative (COR), Federal Task Lead, project teams, vendors, and other stakeholders, the Change Control & Configuration Manager manages resources, risks, dependencies, and customer relationships and supports successful delivery in accordance with task order requirements, federal standards, and DOI and IA policies and procedures. DUTIES AND RESPONSIBILITIES Change & Configuration Management Maintain the OIT Configuration Management Plan and associated processes, procedures, standards, templates, and supporting documentation in accordance with applicable NIST, DOI, IA, and OIT requirements. Support configuration management activities using the IA Configuration Management Database (CMDB), approved repositories, enterprise systems, and authorized manual processes. Support configuration identification, configuration control, status accounting, verification, and audit activities. Ensure Configuration Items (CIs) and associated records remain current, complete, accurate, traceable, and appropriately maintained. Maintain and validate configuration baselines, records, documentation, and supporting artifacts. Coordinate change-control activities from request intake and analysis through review, approval, implementation, validation, and closure. Coordinate release-management activities using available automated and manual tools and processes. Support implementation and continuous improvement of configuration management services, concepts, tools, policies, standards, processes, and procedures. Facilitate Configuration Control Board (CCB) meetings, including meeting preparation, agenda development, change documentation, decision tracking, action items, approvals, and follow-up activities. Document CCB decisions and coordinate implementation and tracking of approved changes. Evaluate proposed changes for potential impacts to scope, schedule, cost, security, operations, configuration items, dependencies, and project outcomes. Coordinate configuration-management audits and support responses to internal and external audit requests. Track audit findings, corrective actions, remediation activities, and associated documentation through resolution. Analyze configuration, change, and release-management processes and recommend opportunities to improve efficiency, data quality, traceability, governance, compliance, and service delivery. Develop and maintain CCM reports, dashboards, metrics, workflows, process models, procedures, and other management artifacts. Coordinate with IT Asset Management, Quality Management, security, service management, business analysis, and technical teams to maintain alignment between configuration records and operational activities. Project Management Manage assigned IT projects throughout the project lifecycle, including initiation, planning, execution, monitoring and control, deployment, transition, and closeout. Apply Project Management Institute (PMI) principles, Project Management Body of Knowledge (PMBOK) practices, Agile methodologies, Waterfall methodologies, and hybrid approaches appropriate to project requirements. Lead and facilitate Program Increment (PI) Planning activities and associated planning, coordination, and follow-up. Develop and maintain project plans, schedules, roadmaps, guides, templates, forms, governance artifacts, decision records, status reports, and other required project documentation. Coordinate information collection, review, approval, version control, and maintenance of project records. Facilitate Agile ceremonies and maintain project and sprint backlogs in alignment with organizational priorities. Coordinate personnel, vendors, technical teams, Government stakeholders, and other resources required to achieve project objectives. Establish and track milestones, dependencies, deliverables, action items, and performance requirements. Monitor project expenditures, compare actual costs against approved budgets, develop projections, and analyze cost, schedule, and performance requirements. Identify, assess, document, track, mitigate, and resolve risks, issues, dependencies, and impediments. Develop mitigation and contingency plans and escalate significant risks and issues in a timely manner. Collaborate with technical teams, vendors, CORs, Federal Task Leads, and other Government stakeholders to support successful project delivery and contract compliance. Maintain project plans, status reports, decision records, test plans, governance artifacts, dashboards, and other required documentation. Conduct stakeholder analysis and develop project communications plans that support project objectives, decision-making, organizational readiness, and stakeholder coordination. Monitor project performance and recommend corrective actions when scope, schedule, cost, quality, risk, compliance, or delivery requirements are at risk. Functional & Delivery Integration Apply Change and Configuration Management expertise within an integrated project-management environment. Provide substantive functional support in at least one complementary discipline: Quality & Test Management IT Asset & Software License Management Business Analysis & Requirements Management Integrate change and configuration activities with applicable project requirements, schedules, testing, releases, assets, risks, security requirements, and implementation activities. Coordinate across Government, business, technical, security, vendor, and contractor teams to support requirements, delivery, risk, quality, change, documentation, and project outcomes. Identify dependencies between configuration items, approved changes, releases, requirements, assets, testing activities, and project deliverables. Support modernization and service-delivery improvements by applying consistent configuration, change-control, governance, and project-management practices. Qualifications REQUIRED EXPERIENCE AND QUALIFICATIONS Minimum of five (5) years of project management experience. Current Project Management Professional (PMP) certification. Demonstrated experience with Change Control, Configuration Management, and/or Release Management in an enterprise IT environment. Experience supporting configuration-management plans, configuration identification, status accounting, configuration records, CCB processes, change control, release management, and configuration audits. Experience working with Configuration Management Databases (CMDBs), configuration repositories, or comparable enterprise configuration-management systems. Strong experience with Agile and Waterfall project management methodologies. Hands-on experience leading Program Increment (PI) Planning activities. In-depth knowledge of IT project management principles and practices, including those established by PMI and PMBOK. Demonstrated hands-on experience in at least one complementary IT delivery discipline: . click apply for full job details
09/23/2026
Full time
Job Description Job Description Company Description Big challenges need bold thinkers. If you're someone who sees problems as opportunities, you'll thrive here. RESPEC is 100% employee-owned, which means we take ownership of every challenge. Here, your ideas drive real solutions. Since 1969, we've tackled complex challenges in energy transition, infrastructure resilience, digital transformation, and sustainability. At RESPEC, you'll work alongside clients to take on critical problems. Depending on your expertise, you might design infrastructure in remote locations, develop renewable energy solutions for global projects, or apply data-driven technology to improve mining and water systems. We bring deep technical knowledge, real-world experience, and a commitment to work that matters. If you're looking for a place where your contributions have real impact, you'll fit right in. We do not accept unsolicited resumes from third-party recruiters. Job Description RESPEC is looking to hire an experienced Change Control & Configuration Manager - Intermediate to support Indian Affairs (IA) and Office of Information Technology (OIT) programs, projects, and operational initiatives. The Change Control & Configuration Manager provides specialized expertise in configuration management, change control, and release management while managing assigned IT projects throughout the project lifecycle. This position serves as an intermediate-level Project Manager and coordinates Government, business, technical, security, vendor, and contractor stakeholders to manage scope, schedule, cost, quality, risk, configuration, change, documentation, compliance, and delivery performance. The position is responsible for maintaining effective Change and Configuration Management (CCM) processes, ensuring configuration information and records remain accurate and current, coordinating change and release activities, supporting Configuration Control Board (CCB) activities, and helping ensure compliance with applicable NIST, Department of the Interior (DOI), Indian Affairs (IA), and OIT requirements. In addition to Change and Configuration Management expertise, the successful candidate must bring substantive experience in at least one complementary IT delivery discipline: Quality & Test Management, IT Asset Management, or Business Analysis & Requirements Management. Working with the Contracting Officer's Representative (COR), Federal Task Lead, project teams, vendors, and other stakeholders, the Change Control & Configuration Manager manages resources, risks, dependencies, and customer relationships and supports successful delivery in accordance with task order requirements, federal standards, and DOI and IA policies and procedures. DUTIES AND RESPONSIBILITIES Change & Configuration Management Maintain the OIT Configuration Management Plan and associated processes, procedures, standards, templates, and supporting documentation in accordance with applicable NIST, DOI, IA, and OIT requirements. Support configuration management activities using the IA Configuration Management Database (CMDB), approved repositories, enterprise systems, and authorized manual processes. Support configuration identification, configuration control, status accounting, verification, and audit activities. Ensure Configuration Items (CIs) and associated records remain current, complete, accurate, traceable, and appropriately maintained. Maintain and validate configuration baselines, records, documentation, and supporting artifacts. Coordinate change-control activities from request intake and analysis through review, approval, implementation, validation, and closure. Coordinate release-management activities using available automated and manual tools and processes. Support implementation and continuous improvement of configuration management services, concepts, tools, policies, standards, processes, and procedures. Facilitate Configuration Control Board (CCB) meetings, including meeting preparation, agenda development, change documentation, decision tracking, action items, approvals, and follow-up activities. Document CCB decisions and coordinate implementation and tracking of approved changes. Evaluate proposed changes for potential impacts to scope, schedule, cost, security, operations, configuration items, dependencies, and project outcomes. Coordinate configuration-management audits and support responses to internal and external audit requests. Track audit findings, corrective actions, remediation activities, and associated documentation through resolution. Analyze configuration, change, and release-management processes and recommend opportunities to improve efficiency, data quality, traceability, governance, compliance, and service delivery. Develop and maintain CCM reports, dashboards, metrics, workflows, process models, procedures, and other management artifacts. Coordinate with IT Asset Management, Quality Management, security, service management, business analysis, and technical teams to maintain alignment between configuration records and operational activities. Project Management Manage assigned IT projects throughout the project lifecycle, including initiation, planning, execution, monitoring and control, deployment, transition, and closeout. Apply Project Management Institute (PMI) principles, Project Management Body of Knowledge (PMBOK) practices, Agile methodologies, Waterfall methodologies, and hybrid approaches appropriate to project requirements. Lead and facilitate Program Increment (PI) Planning activities and associated planning, coordination, and follow-up. Develop and maintain project plans, schedules, roadmaps, guides, templates, forms, governance artifacts, decision records, status reports, and other required project documentation. Coordinate information collection, review, approval, version control, and maintenance of project records. Facilitate Agile ceremonies and maintain project and sprint backlogs in alignment with organizational priorities. Coordinate personnel, vendors, technical teams, Government stakeholders, and other resources required to achieve project objectives. Establish and track milestones, dependencies, deliverables, action items, and performance requirements. Monitor project expenditures, compare actual costs against approved budgets, develop projections, and analyze cost, schedule, and performance requirements. Identify, assess, document, track, mitigate, and resolve risks, issues, dependencies, and impediments. Develop mitigation and contingency plans and escalate significant risks and issues in a timely manner. Collaborate with technical teams, vendors, CORs, Federal Task Leads, and other Government stakeholders to support successful project delivery and contract compliance. Maintain project plans, status reports, decision records, test plans, governance artifacts, dashboards, and other required documentation. Conduct stakeholder analysis and develop project communications plans that support project objectives, decision-making, organizational readiness, and stakeholder coordination. Monitor project performance and recommend corrective actions when scope, schedule, cost, quality, risk, compliance, or delivery requirements are at risk. Functional & Delivery Integration Apply Change and Configuration Management expertise within an integrated project-management environment. Provide substantive functional support in at least one complementary discipline: Quality & Test Management IT Asset & Software License Management Business Analysis & Requirements Management Integrate change and configuration activities with applicable project requirements, schedules, testing, releases, assets, risks, security requirements, and implementation activities. Coordinate across Government, business, technical, security, vendor, and contractor teams to support requirements, delivery, risk, quality, change, documentation, and project outcomes. Identify dependencies between configuration items, approved changes, releases, requirements, assets, testing activities, and project deliverables. Support modernization and service-delivery improvements by applying consistent configuration, change-control, governance, and project-management practices. Qualifications REQUIRED EXPERIENCE AND QUALIFICATIONS Minimum of five (5) years of project management experience. Current Project Management Professional (PMP) certification. Demonstrated experience with Change Control, Configuration Management, and/or Release Management in an enterprise IT environment. Experience supporting configuration-management plans, configuration identification, status accounting, configuration records, CCB processes, change control, release management, and configuration audits. Experience working with Configuration Management Databases (CMDBs), configuration repositories, or comparable enterprise configuration-management systems. Strong experience with Agile and Waterfall project management methodologies. Hands-on experience leading Program Increment (PI) Planning activities. In-depth knowledge of IT project management principles and practices, including those established by PMI and PMBOK. Demonstrated hands-on experience in at least one complementary IT delivery discipline: . click apply for full job details
IT/OT Technician - KS, MO, NE, & AR
The Industrial Solutions Network of CED Omaha, Nebraska
Job DescriptionJob DescriptionAIMM Services is a specialized team dedicated to providing expert services and assessments to the manufacturing industry. As part of the Industrial Solutions Network, AIMM Services supports U.S. manufacturing businesses with solutions that enhance competitiveness and drive success. Our collaborative culture fosters both personal and professional growth, making AIMM an exciting place to build your career. Recruiting for: AR, KS, MO, & NE Position summary: Are you passionate about bridging the gap between Information Technology (IT) and Operational Technology (OT) in industrial environments? As an IT/OT Technician at AIMM Services, you'll play a key role in assessing, designing, and implementing enterprise-level software and hardware solutions for industrial and manufacturing clients. This role focuses on network security, infrastructure stability, and digital transformation, ensuring clients operate with secure, reliable, and scalable IT/OT systems. You'll work closely with clients in a consultative, solution-driven capacity, helping them optimize their technology landscape while collaborating with industry leaders such as Rockwell Automation, Cisco, Dell, Microsoft, and VMware. What you'll do: Technical Solutions: Implement and support solutions from leading technology partners (Rockwell, Cisco, Dell, Microsoft, etc.). Configure and deploy servers, network switches, firewalls, VPNs, and virtualized environments. Conduct network and cybersecurity assessments to identify risks and recommend solutions. Lead the deployment of IT/OT solutions, including FactoryTalk Services, ThinManager, VMWare, and AssetCentre. Develop standard operating procedures for IT and security offerings. Document findings and present them to technical and non-technical stakeholders. Client-Focused Solutions: Act as a trusted advisor, helping clients understand their IT/OT infrastructure and implement improvements. Collaborate with clients' IT and OT teams to ensure seamless integration of technologies. Provide technical education, training, and webinars to internal teams and clients. Support sales teams with expert insights into IT, security, and network solutions. What we're looking for: Bachelor's Degree in Electrical Engineering, Industrial Engineering, Computer Science, or IT (or equivalent experience). 3+ years of experience in network engineering, IT infrastructure, or cybersecurity. Hands-on expertise with network hardware, virtualization, and industrial control systems. Proficiency in configuring routers, switches, firewalls, servers, VPNs, and Microsoft environments. Industry certifications such as CCENT, CCNA, VCA, VCP, GICSP (preferred). Strong understanding of ISA, ANSI, NEMA, NEC, NIST standards (preferred). Excellent problem-solving, analytical, and communication skills. Travel Requirements: 50% - 75% travel to meet clients and oversee implementations. If you're ready to drive innovation at the intersection of IT and OT, apply today and take the next step in your career! The Industrial Solutions Network is part of Consolidated Electrical Distributors, CED Inc. CED is an Equal Opportunity Employer/Disability and Veteran Status. For more information visit: Please Note: This is NOT the official application for this position. The official application will be sent later in the interview process. NOTE: This job description is not designed to cover or contain a comprehensive listing of all required activities, duties or responsibilities. Other duties, responsibilities, and activities may be assigned at any time; with or without notice. We are an Equal Opportunity Employer - Disability Veteran All references to Company/We mean CONSOLIDATED ELECTRICAL DISTRIBUTORS Powered by JazzHR aict3t0b3Q
09/23/2026
Full time
Job DescriptionJob DescriptionAIMM Services is a specialized team dedicated to providing expert services and assessments to the manufacturing industry. As part of the Industrial Solutions Network, AIMM Services supports U.S. manufacturing businesses with solutions that enhance competitiveness and drive success. Our collaborative culture fosters both personal and professional growth, making AIMM an exciting place to build your career. Recruiting for: AR, KS, MO, & NE Position summary: Are you passionate about bridging the gap between Information Technology (IT) and Operational Technology (OT) in industrial environments? As an IT/OT Technician at AIMM Services, you'll play a key role in assessing, designing, and implementing enterprise-level software and hardware solutions for industrial and manufacturing clients. This role focuses on network security, infrastructure stability, and digital transformation, ensuring clients operate with secure, reliable, and scalable IT/OT systems. You'll work closely with clients in a consultative, solution-driven capacity, helping them optimize their technology landscape while collaborating with industry leaders such as Rockwell Automation, Cisco, Dell, Microsoft, and VMware. What you'll do: Technical Solutions: Implement and support solutions from leading technology partners (Rockwell, Cisco, Dell, Microsoft, etc.). Configure and deploy servers, network switches, firewalls, VPNs, and virtualized environments. Conduct network and cybersecurity assessments to identify risks and recommend solutions. Lead the deployment of IT/OT solutions, including FactoryTalk Services, ThinManager, VMWare, and AssetCentre. Develop standard operating procedures for IT and security offerings. Document findings and present them to technical and non-technical stakeholders. Client-Focused Solutions: Act as a trusted advisor, helping clients understand their IT/OT infrastructure and implement improvements. Collaborate with clients' IT and OT teams to ensure seamless integration of technologies. Provide technical education, training, and webinars to internal teams and clients. Support sales teams with expert insights into IT, security, and network solutions. What we're looking for: Bachelor's Degree in Electrical Engineering, Industrial Engineering, Computer Science, or IT (or equivalent experience). 3+ years of experience in network engineering, IT infrastructure, or cybersecurity. Hands-on expertise with network hardware, virtualization, and industrial control systems. Proficiency in configuring routers, switches, firewalls, servers, VPNs, and Microsoft environments. Industry certifications such as CCENT, CCNA, VCA, VCP, GICSP (preferred). Strong understanding of ISA, ANSI, NEMA, NEC, NIST standards (preferred). Excellent problem-solving, analytical, and communication skills. Travel Requirements: 50% - 75% travel to meet clients and oversee implementations. If you're ready to drive innovation at the intersection of IT and OT, apply today and take the next step in your career! The Industrial Solutions Network is part of Consolidated Electrical Distributors, CED Inc. CED is an Equal Opportunity Employer/Disability and Veteran Status. For more information visit: Please Note: This is NOT the official application for this position. The official application will be sent later in the interview process. NOTE: This job description is not designed to cover or contain a comprehensive listing of all required activities, duties or responsibilities. Other duties, responsibilities, and activities may be assigned at any time; with or without notice. We are an Equal Opportunity Employer - Disability Veteran All references to Company/We mean CONSOLIDATED ELECTRICAL DISTRIBUTORS Powered by JazzHR aict3t0b3Q
EY
Service Delivery Center - Apigee Migration Engineer - Senior
EY Dallas, Texas
Location: San Antonio Farinon Park, Dallas - 1201 Elm St At EY, we're all in to shape your future with confidence. We'll help you succeed in a globally connected powerhouse of diverse teams and take your career wherever you want it to go. Join EY and help to build a better working world. The opportunity We are offering you a demanding role. You will use the most advanced Quality models to implement the newest delivery excellence solutions. This position includes possibility to interact within international environment and work with the most recognizable and influential players in Delivery Excellence space. Your key responsibilities Responsible for designing, developing, implementing, and supporting enterprise API solutions using Apigee, ensuring secure, scalable, and high-performing API ecosystems. Accountable for API lifecycle management including API design, development, security, deployment, monitoring, versioning, and governance. Assess, design, build, test, deploy, and document API integrations between enterprise applications, cloud platforms, third-party vendors, and partner systems. Design and implement reusable API proxies, shared flows, security policies, and common integration components. Must be comfortable working in Agile methodology, driving technical discussions, gathering requirements from multiple stakeholders, and preparing technical solution specifications. Develop API management solutions including traffic management, authentication, authorization, rate limiting, caching, and threat protection. Identify root causes of production issues and provide technical solutions for performance optimization and operational stability. Forecast technical risks, dependencies, and mitigation plans and communicate them proactively with architects and delivery managers. Ensure API governance standards, security compliance, and enterprise integration best practices are followed across projects. Contribute to organization assets, accelerators, frameworks, reusable components, and process improvements. Deliver Proof of Concepts (PoCs), conduct code reviews, and mentor junior developers. Implement DevOps and CI/CD practices for API development, testing, deployment, and monitoring. Support integration testing, performance testing, security testing, and user acceptance testing activities. Collaborate with business, security, infrastructure, and application teams to deliver scalable API-led connectivity solutions. Key skills: Experience in API Management and Integration projects and hands-on experience in Apigee development. Must have implementation experience using Apigee in at least 2 enterprise-scale projects. Strong experience with Apigee Edge and/or Apigee X. Experience designing and implementing RESTful APIs and API-first architectures. Strong hands-on experience with API Proxy development, Shared Flows, Flow Hooks, API Products, Developers, Apps, and Monetization concepts. Experience implementing API security standards including OAuth 2.0, OpenID Connect, JWT, SAML, API Keys, mTLS, and SSL/TLS. Experience with API Gateway policies such as Spike Arrest, Quota, Caching, Threat Protection, Traffic Management, Message Transformation, and Logging. Strong knowledge of JavaScript and policy-based API development within Apigee. Experience integrating enterprise applications, ERP systems, CRM applications, SaaS platforms, and cloud-native services. Experience with REST, SOAP, JSON, XML, OpenAPI/Swagger specifications, Microservices, and Event-Driven Architecture. Hands-on experience troubleshooting API performance, latency, scalability, and security issues. Experience with logging and monitoring tools such as Splunk, ELK, Dynatrace, Grafana, Cloud Monitoring, or equivalent platforms. Knowledge of CI/CD tools including Jenkins, GitHub, GitLab, Azure DevOps, Maven, and automated deployment pipelines. Experience working with Kubernetes, Docker, and cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Understanding of API Governance, API Analytics, Developer Portals, and API Lifecycle Management. Multi-domain expertise and experience with other integration platforms is an added advantage. Strong written and verbal communication skills. Apigee API Engineer Certification or related Google Cloud certifications are preferred. Should be capable of mentoring team members and handling customer interactions independently. Skills and attributes for success: Strong communication skills and ability to collaborate with business delivery teams. Analytical thinking. Strong team-player and self-started attitude. Ability to create quality process documentation. Well organized, attentive to details attitude. To qualify for the role, you must have: College degree in Computer Science, Engineering, Information Technology, or related field. Minimum 3 years of experience in API development, integration, and API management projects. Strong understanding of API Security, API Governance, and Enterprise Integration Patterns. Experience handling production support, incident management, and root cause analysis. Ability to manage technical risks and customer escalations. Ability to interact effectively across business, technical, and leadership teams. Experience working in Agile Scrum and DevOps environments. Additionally, good to have: Experience with other API Management platforms such as IBM API Connect (APIC) and DataPower will be an added advantage. Experience with event streaming platforms such as Kafka or Pub/Sub. Knowledge of project management tools like JIRA, Azure DevOps, and Confluence. Exposure to DevOps automation, Infrastructure as Code (Terraform), and cloud-native architectures. Knowledge of SDLC, Waterfall, Agile, SAFe Agile, and Scrum methodologies. Experience with API monetization and partner onboarding solutions. Understanding of enterprise architecture frameworks and digital transformation programs. Ideally, you'll also have: Ability to define processes for teams. Experience in project audits and reviews. Ability to analyse risks and metrics of projects. What we look for: A Team of people with commercial acumen, technical experience and enthusiasm to learn new things in this fast-moving environment An opportunity to be a part of market-leading, multi-disciplinary team of 250+ professionals, in the only integrated global transaction business worldwide. Opportunities to work with EY ServiceNow practices globally with leading businesses across a range of industries What working at EY offers At EY, we're dedicated to helping our clients, from start-ups to Fortune 500 companies - and the work we do with them is as varied as they are. You get to work with inspiring and meaningful projects. Our focus is education and coaching alongside practical experience to ensure your personal development. We value our employees and you will be able to control your own development with an individual progression plan. You will quickly grow into a responsible role with challenging and stimulating assignments. Moreover, you will be part of an interdisciplinary environment that emphasizes high quality and knowledge exchange. Plus, we offer: Support, coaching and feedback from some of the most engaging colleagues around Opportunities to develop new skills and progress your career The freedom and flexibility to handle your role in a way that's right for you About EY As a global leader in assurance, tax, transaction and advisory services, we're using the finance products, expertise and systems we've developed to build a better working world. That starts with a culture that believes in giving you the training, opportunities and creative freedom to make things better. Whenever you join, however long you stay, the exceptional EY experience lasts a lifetime. And with a commitment to hiring and developing the most passionate people, we'll make our ambition to be the best employer by 2020 a reality. If you can confidently demonstrate that you meet the criteria above, please contact us as soon as possible. Join us in building a better working world. Apply now What we offer you At EY, we harness our collective strength to empower you to shape your future with confidence through professional growth, personal fulfillment and an inclusive culture. Learn more at We offer a comprehensive compensation and benefits package where you'll be rewarded based on your performance and recognized for the value you bring to the business. The base salary range for this job is: New York City, Boston, and Washington DC Metro Areas, Washington State, and Southern California offices - $80,300 to $149,100 Bay Area California offices - $83,700 to $155,300 All other offices locations in the US, including Sacramento - $67,000 to $136,800 Individual salaries within these ranges are determined through a wide variety of factors including but not limited to education, experience, knowledge, skills and geography. In addition, our Total Rewards package includes medical and dental coverage . click apply for full job details
09/23/2026
Full time
Location: San Antonio Farinon Park, Dallas - 1201 Elm St At EY, we're all in to shape your future with confidence. We'll help you succeed in a globally connected powerhouse of diverse teams and take your career wherever you want it to go. Join EY and help to build a better working world. The opportunity We are offering you a demanding role. You will use the most advanced Quality models to implement the newest delivery excellence solutions. This position includes possibility to interact within international environment and work with the most recognizable and influential players in Delivery Excellence space. Your key responsibilities Responsible for designing, developing, implementing, and supporting enterprise API solutions using Apigee, ensuring secure, scalable, and high-performing API ecosystems. Accountable for API lifecycle management including API design, development, security, deployment, monitoring, versioning, and governance. Assess, design, build, test, deploy, and document API integrations between enterprise applications, cloud platforms, third-party vendors, and partner systems. Design and implement reusable API proxies, shared flows, security policies, and common integration components. Must be comfortable working in Agile methodology, driving technical discussions, gathering requirements from multiple stakeholders, and preparing technical solution specifications. Develop API management solutions including traffic management, authentication, authorization, rate limiting, caching, and threat protection. Identify root causes of production issues and provide technical solutions for performance optimization and operational stability. Forecast technical risks, dependencies, and mitigation plans and communicate them proactively with architects and delivery managers. Ensure API governance standards, security compliance, and enterprise integration best practices are followed across projects. Contribute to organization assets, accelerators, frameworks, reusable components, and process improvements. Deliver Proof of Concepts (PoCs), conduct code reviews, and mentor junior developers. Implement DevOps and CI/CD practices for API development, testing, deployment, and monitoring. Support integration testing, performance testing, security testing, and user acceptance testing activities. Collaborate with business, security, infrastructure, and application teams to deliver scalable API-led connectivity solutions. Key skills: Experience in API Management and Integration projects and hands-on experience in Apigee development. Must have implementation experience using Apigee in at least 2 enterprise-scale projects. Strong experience with Apigee Edge and/or Apigee X. Experience designing and implementing RESTful APIs and API-first architectures. Strong hands-on experience with API Proxy development, Shared Flows, Flow Hooks, API Products, Developers, Apps, and Monetization concepts. Experience implementing API security standards including OAuth 2.0, OpenID Connect, JWT, SAML, API Keys, mTLS, and SSL/TLS. Experience with API Gateway policies such as Spike Arrest, Quota, Caching, Threat Protection, Traffic Management, Message Transformation, and Logging. Strong knowledge of JavaScript and policy-based API development within Apigee. Experience integrating enterprise applications, ERP systems, CRM applications, SaaS platforms, and cloud-native services. Experience with REST, SOAP, JSON, XML, OpenAPI/Swagger specifications, Microservices, and Event-Driven Architecture. Hands-on experience troubleshooting API performance, latency, scalability, and security issues. Experience with logging and monitoring tools such as Splunk, ELK, Dynatrace, Grafana, Cloud Monitoring, or equivalent platforms. Knowledge of CI/CD tools including Jenkins, GitHub, GitLab, Azure DevOps, Maven, and automated deployment pipelines. Experience working with Kubernetes, Docker, and cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Understanding of API Governance, API Analytics, Developer Portals, and API Lifecycle Management. Multi-domain expertise and experience with other integration platforms is an added advantage. Strong written and verbal communication skills. Apigee API Engineer Certification or related Google Cloud certifications are preferred. Should be capable of mentoring team members and handling customer interactions independently. Skills and attributes for success: Strong communication skills and ability to collaborate with business delivery teams. Analytical thinking. Strong team-player and self-started attitude. Ability to create quality process documentation. Well organized, attentive to details attitude. To qualify for the role, you must have: College degree in Computer Science, Engineering, Information Technology, or related field. Minimum 3 years of experience in API development, integration, and API management projects. Strong understanding of API Security, API Governance, and Enterprise Integration Patterns. Experience handling production support, incident management, and root cause analysis. Ability to manage technical risks and customer escalations. Ability to interact effectively across business, technical, and leadership teams. Experience working in Agile Scrum and DevOps environments. Additionally, good to have: Experience with other API Management platforms such as IBM API Connect (APIC) and DataPower will be an added advantage. Experience with event streaming platforms such as Kafka or Pub/Sub. Knowledge of project management tools like JIRA, Azure DevOps, and Confluence. Exposure to DevOps automation, Infrastructure as Code (Terraform), and cloud-native architectures. Knowledge of SDLC, Waterfall, Agile, SAFe Agile, and Scrum methodologies. Experience with API monetization and partner onboarding solutions. Understanding of enterprise architecture frameworks and digital transformation programs. Ideally, you'll also have: Ability to define processes for teams. Experience in project audits and reviews. Ability to analyse risks and metrics of projects. What we look for: A Team of people with commercial acumen, technical experience and enthusiasm to learn new things in this fast-moving environment An opportunity to be a part of market-leading, multi-disciplinary team of 250+ professionals, in the only integrated global transaction business worldwide. Opportunities to work with EY ServiceNow practices globally with leading businesses across a range of industries What working at EY offers At EY, we're dedicated to helping our clients, from start-ups to Fortune 500 companies - and the work we do with them is as varied as they are. You get to work with inspiring and meaningful projects. Our focus is education and coaching alongside practical experience to ensure your personal development. We value our employees and you will be able to control your own development with an individual progression plan. You will quickly grow into a responsible role with challenging and stimulating assignments. Moreover, you will be part of an interdisciplinary environment that emphasizes high quality and knowledge exchange. Plus, we offer: Support, coaching and feedback from some of the most engaging colleagues around Opportunities to develop new skills and progress your career The freedom and flexibility to handle your role in a way that's right for you About EY As a global leader in assurance, tax, transaction and advisory services, we're using the finance products, expertise and systems we've developed to build a better working world. That starts with a culture that believes in giving you the training, opportunities and creative freedom to make things better. Whenever you join, however long you stay, the exceptional EY experience lasts a lifetime. And with a commitment to hiring and developing the most passionate people, we'll make our ambition to be the best employer by 2020 a reality. If you can confidently demonstrate that you meet the criteria above, please contact us as soon as possible. Join us in building a better working world. Apply now What we offer you At EY, we harness our collective strength to empower you to shape your future with confidence through professional growth, personal fulfillment and an inclusive culture. Learn more at We offer a comprehensive compensation and benefits package where you'll be rewarded based on your performance and recognized for the value you bring to the business. The base salary range for this job is: New York City, Boston, and Washington DC Metro Areas, Washington State, and Southern California offices - $80,300 to $149,100 Bay Area California offices - $83,700 to $155,300 All other offices locations in the US, including Sacramento - $67,000 to $136,800 Individual salaries within these ranges are determined through a wide variety of factors including but not limited to education, experience, knowledge, skills and geography. In addition, our Total Rewards package includes medical and dental coverage . click apply for full job details
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler Short Hills, New Jersey
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler San Jose, California
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler McLean, Virginia
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Staff Site Reliability Engineer (Production Engineer)- Federal
Zscaler Boston, Massachusetts
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
09/23/2026
Full time
Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange ️ platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world's largest in-line cloud security platform. We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world's hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler. Role We are looking for a Staff Site Reliability Engineer (Production Engineer) to join our team. This is a hybrid role (onsite three days a week in San Jose, CA or another Zscaler office; remote can be considered for exceptional candidates) reporting to the Senior Manager, Site Reliability Engineering in the Zero Trust Exchange department. As a key member of the Zero Trust Exchange team, you will own the systems-level reliability and performance of Zscaler's high-throughput bare-metal and cloud infrastructure processing tens of billions of daily transactions across a global, multi-region fleet. This is a software-first SRE role: you will write production-grade code and automation, drive the shift from reactive incident response, and bring engineering discipline to the systems-level work - OS, network and application debugging - that keeps the fleet operating safely at scale. What You'll Do (Role Expectations) Maintain high availability across large-scale bare-metal Linux/BSD fleets, Kubernetes clusters, and custom routing stacks in partnership with Engineering and Networking teams Lead full-cycle incident response by conducting cross-stack troubleshooting using low-level OS and network tools (strace, lsof, tcpdump, iostat, vmstat, gdb), maintain high availability across large-scale bare metal Linux /BSD fleets and Kubernetes clusters Automate infrastructure lifecycle management, service provisioning, configuration workflows, and release deployments using Ansible, Python, and Go; quantify operational toil and convert recurring manual work into durable, version-controlled, testable automation - tracking reduction as an engineering outcome Own end-to-end telemetry (metrics, logs, traces) using Prometheus and OpenTelemetry ecosystems; define and enforce SLOs/error budgets to reduce alert noise Perform architectural reviews, OS/kernel upgrades, capacity and performance tuning, strict CI/CD validation prior to production rollouts; embed operability standards (telemetry, rollback safety, SLO readiness) into service design from the start Who You Are (Success Profile) You thrive in ambiguity. You're comfortable building the path as you walk it, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful. You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution. You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact. You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust. You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose. What We're Looking For (Minimum Qualifications) US Citizenship is required (due to the nature of assigned customers) Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain 5+ years of experience in Site Reliability Engineering, Production Engineering, or Systems Engineering operating high-scale, low-latency production platforms Proven ability to write and debug executable code live (Python, Go, or Bash) covering core logic/data structures, along with hands-on experience writing Ansible playbooks/tasks for infrastructure automation Deep knowledge of Linux OS internals and kernel troubleshooting (e.g., inodes, open file descriptors, process states, and analyzing df vs du storage discrepancies) Comprehensive understanding of networking protocols and packet-level analysis, including DNS resolution workflows, TLS handshakes, TCP/IP mechanics, and packet captures via tcpdump What Will Make You Stand Out (Preferred Qualifications) Hands-on experience operating and managing FreeBSD / BSD operating systems in production Proven expertise running, scaling, and troubleshooting Kubernetes clusters in high-traffic, low-latency environments and workflow orchestration platforms (Temporal or similar) Deep experience with Prometheus / OpenTelemetry ecosystems, or leveraging AI/ML frameworks/AIOps tools for automated root-cause analysis Zscaler's salary ranges are benchmarked and are determined by role and level. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position across all US locations and could be higher or lower based on a multitude of factors, including job-related skills, experience, and relevant education or training. The base salary range listed for this full-time position excludes commission/ bonus/ equity (if applicable) + benefits. Base Pay Range $119,000-$170,000 USD At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure. Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including: Various health plans Time off plans for vacation and sick time Parental leave options Retirement options Education reimbursement In-office perks, and more! Learn more about Zscaler's hybrid working model and benefits here. By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines. Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link. Pay Transparency Zscaler complies with all applicable federal, state, and local pay transparency rules. Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.
Software Engineer II
Cross River Fort Lee, New Jersey
Who We Are Cross River builds the infrastructure behind the world's most innovative financial products. Our technology and capital solutions power payments, cards, lending, and digital asset capabilities that move money safely, instantly, and inclusively - trusted by leading fintechs, enterprises, and disruptors across the globe. Our mission is simple: to build the financial infrastructure that expands access and opportunity for all. Guided by a culture of collaboration, curiosity, and purpose, Cross River has been named one of American Banker's Best Places to Work in Fintech year after year. Whether you're designing code, solving regulatory puzzles, or developing strategy, you'll join a team where innovation and integrity drive everything we do - and where your work helps shape the future of finance. About Our Team Cross River's AI Transformation team is a newly formed group of problem solvers passionate about bridging the gap between cutting-edge AI capabilities and real-world business needs. Based in-office at our Fort Lee, New Jersey headquarters, we work face-to-face with stakeholders across the organization to identify opportunities, adopt and adapt AI tools, and build sustainable solutions. Our mission is to bring AI-powered solutions to life, ensure they are maintainable, auditable, and resilient, and to continuously leverage advances in AI technologies for the benefit of the business. What We're Looking For We are looking for a Software Engineer II to focus on solutions for our AI Transformation team. You will work in-office at least four days per week at our Fort Lee, New Jersey headquarters. With guidance from your Engineering Manager and a Senior Software Engineer, you will collaborate closely with an AI strategy team and collaborate face-to-face with business stakeholders to deeply understand their problems, cut through surface-level descriptions to identify root challenges, and write the code that solves real problems. You will work closely with other software and QA engineers in and outside of your team. We are looking for someone who has worked within ambiguity, and seeks out and applies strong engineering disciplines to deliver sustainable, AI-centered or augmented outcomes. Responsibilities: Design, develop, and deliver scalable software modules and components Build backend systems using a mix of technologies. Languages may include C# (.NET 10), Python, and JavaScript/TypeScript. Databases may include PostgreSQL and SQL Server. Infrastructure includes AWS, Docker, and Kubernetes Implement, test, and iterate on software solutions that address both short-term and long-term considerations with a maintainable, sustainable mindset Qualifications: 4+ years of experience developing enterprise systems in modern, current object-oriented languages, with at least 2+ years of C# 2+ years of experience with SQL, preferably PostgreSQL or SQL Server Experience working with AI assistive and agentic code generation (e.g. Copilot, Claude Code, Cursor) Database persistence frameworks (e.g. nHibernate, Entity Framework) Strong communication skills Experience or understanding of Domain Driven Design Cloud Architecture - preferably AWS Docker / Containers Financial industry / accounting experience or understanding is helpful, but not required Experience designing and developing distributed systems and event driven architecture is preferred. Ideally with understanding or exposure to: MassTransit or NServiceBus RabbitMQ Idempotency Salary Range: $130,000.00 - $170,000.00 Cross River is an Equal Opportunity Employer. Cross River does not discriminate on the basis of race, religion, color, sex, gender identity, sexual orientation, age, non-disqualifying physical or mental disability, national origin, veteran status or any other basis covered by appropriate law. All employment is decided on the basis of qualifications, merit, and business need. By submitting your application, you give Cross River permission to email, call, or text you using the contact details provided. We will only contact you with job related information.
09/23/2026
Full time
Who We Are Cross River builds the infrastructure behind the world's most innovative financial products. Our technology and capital solutions power payments, cards, lending, and digital asset capabilities that move money safely, instantly, and inclusively - trusted by leading fintechs, enterprises, and disruptors across the globe. Our mission is simple: to build the financial infrastructure that expands access and opportunity for all. Guided by a culture of collaboration, curiosity, and purpose, Cross River has been named one of American Banker's Best Places to Work in Fintech year after year. Whether you're designing code, solving regulatory puzzles, or developing strategy, you'll join a team where innovation and integrity drive everything we do - and where your work helps shape the future of finance. About Our Team Cross River's AI Transformation team is a newly formed group of problem solvers passionate about bridging the gap between cutting-edge AI capabilities and real-world business needs. Based in-office at our Fort Lee, New Jersey headquarters, we work face-to-face with stakeholders across the organization to identify opportunities, adopt and adapt AI tools, and build sustainable solutions. Our mission is to bring AI-powered solutions to life, ensure they are maintainable, auditable, and resilient, and to continuously leverage advances in AI technologies for the benefit of the business. What We're Looking For We are looking for a Software Engineer II to focus on solutions for our AI Transformation team. You will work in-office at least four days per week at our Fort Lee, New Jersey headquarters. With guidance from your Engineering Manager and a Senior Software Engineer, you will collaborate closely with an AI strategy team and collaborate face-to-face with business stakeholders to deeply understand their problems, cut through surface-level descriptions to identify root challenges, and write the code that solves real problems. You will work closely with other software and QA engineers in and outside of your team. We are looking for someone who has worked within ambiguity, and seeks out and applies strong engineering disciplines to deliver sustainable, AI-centered or augmented outcomes. Responsibilities: Design, develop, and deliver scalable software modules and components Build backend systems using a mix of technologies. Languages may include C# (.NET 10), Python, and JavaScript/TypeScript. Databases may include PostgreSQL and SQL Server. Infrastructure includes AWS, Docker, and Kubernetes Implement, test, and iterate on software solutions that address both short-term and long-term considerations with a maintainable, sustainable mindset Qualifications: 4+ years of experience developing enterprise systems in modern, current object-oriented languages, with at least 2+ years of C# 2+ years of experience with SQL, preferably PostgreSQL or SQL Server Experience working with AI assistive and agentic code generation (e.g. Copilot, Claude Code, Cursor) Database persistence frameworks (e.g. nHibernate, Entity Framework) Strong communication skills Experience or understanding of Domain Driven Design Cloud Architecture - preferably AWS Docker / Containers Financial industry / accounting experience or understanding is helpful, but not required Experience designing and developing distributed systems and event driven architecture is preferred. Ideally with understanding or exposure to: MassTransit or NServiceBus RabbitMQ Idempotency Salary Range: $130,000.00 - $170,000.00 Cross River is an Equal Opportunity Employer. Cross River does not discriminate on the basis of race, religion, color, sex, gender identity, sexual orientation, age, non-disqualifying physical or mental disability, national origin, veteran status or any other basis covered by appropriate law. All employment is decided on the basis of qualifications, merit, and business need. By submitting your application, you give Cross River permission to email, call, or text you using the contact details provided. We will only contact you with job related information.
Senior DevOps Engineer
Wolters Kluwer Greenwood Village, Colorado
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
09/23/2026
Full time
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
Senior DevOps Engineer
Wolters Kluwer Saint Cloud, Minnesota
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
09/23/2026
Full time
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
Senior DevOps Engineer
Wolters Kluwer Minneapolis, Minnesota
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
09/23/2026
Full time
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
Senior DevOps Engineer
Wolters Kluwer Kennesaw, Georgia
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
09/23/2026
Full time
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
Senior DevOps Engineer
Wolters Kluwer Indianapolis, Indiana
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.
09/23/2026
Full time
As a Senior DevOps Engineer, you will be a key technical contributor responsible for designing, implementing, and operating scalable, resilient infrastructure and CI/CD pipelines that support the full software development lifecycle. You will work closely with Wolters Kluwer Product Teams to embed DevOps best practices into agile workflows, enabling continuous integration, automated testing, and reliable deployment across environments. In this role, you will focus on hands on engineering excellence-building, automating, and operating cloud native platforms that improve system reliability, performance, and maintainability. You will partner closely with application development teams to bridge development and operations, supporting microservices, containerized workloads, and cloud native architectures. While not a formal people manager, you will act as a senior technical mentor and role model, promoting engineering rigor, code as infrastructure, and continuous improvement. Responsibilities Design, engineer, and automate secure, scalable cloud infrastructure in Azure and AWS using Infrastructure as Code (IaC) tools such as Terraform, Ansible, and Jenkins, applying software engineering best practices including modular design, version control, and automated testing. Implement and maintain Infrastructure as Code and CI/CD pipelines, contributing reusable modules, templates, and patterns that improve consistency and reliability across teams. Design and implement modern compute platforms, including containerized and serverless solutions (AKS, EKS, Docker, Azure Functions), with an emphasis on scalability, maintainability, and performance. Build and maintain CI/CD pipelines as software products, ensuring strong test coverage, artifact management, promotion workflows, and deployment automation across multiple environments. Support and evolve cloud native architectures, applying engineering principles such as abstraction, decoupling, fault isolation, and observability. Implement observability solutions using metrics, logging, and tracing to enable proactive issue detection, faster troubleshooting, and root cause analysis. Ensure infrastructure and automation solutions comply with enterprise DevOps, security, and compliance standards, contributing to architectural reviews and governance processes. Serve as a senior technical mentor, providing guidance through code reviews, design discussions, and knowledge sharing-without direct people management responsibilities. Evaluate and prototype emerging tools and technologies, applying engineering rigor to assess value, performance, and integration feasibility. Apply Site Reliability Engineering (SRE) practices such as SLIs/SLOs, error budgets, capacity planning, and incident response to improve system reliability and reduce operational toil. Deploy, operate, and support business critical applications, ensuring high availability, fault tolerance, and performance optimization. Participate in modernization initiatives, supporting the re architecture and cloud native transformation of legacy platforms. Identify and remediate engineering inefficiencies by proposing and implementing automation and architectural improvements. Participate in post incident reviews, contributing to blameless root cause analysis and long term corrective actions. Participate in on call rotations, continuously improving alert quality, reducing noise, and automating remediation where possible. Qualifications Bachelor's degree in Engineering, Computer Science, or a related field (Master's degree preferred). 5+ years of experience in DevOps, Site Reliability Engineering, Release Engineering, or related roles, with strong hands on software engineering experience. Strong background in software development, with experience in languages such as Python, .NET, or Java. Proven experience working with cloud platforms (Azure and AWS). Proficiency in scripting languages such as PowerShell and Bash. Solid understanding of core Azure and AWS services (PaaS, IaaS, SaaS). Strong experience with source control and automation tools, including Git. Hands on experience with Infrastructure as Code tools such as Terraform or CloudFormation. Experience building and operating CI/CD pipelines using tools such as Azure DevOps, Jenkins, or similar platforms. Strong problem solving skills with attention to detail and operational excellence. Ability to clearly communicate technical concepts to engineers and non engineering stakeholders. Demonstrated commitment to DevOps culture, including continuous integration, automated testing, deployment automation, and full lifecycle ownership. Experience building or supporting observability platforms and defining operational best practices. Experience troubleshooting and automating diagnostics across Linux and Windows environments. Our Interview Practices To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. To help maintain this integrity, we ask to remove virtual backgrounds and include in-person interviews in our hiring process. Please note that use of AI-generated responses or third-party support during interviews will be grounds for disqualification from the recruitment process. Applicants may be required to appear onsite at a Wolters Kluwer office as part of the recruitment process. Compensation: $92,700.00 - $161,850.00 USDThis role is eligible for Bonus. Compensation range listed is based on primary location of the position. Actual base salary offer is influenced by a wide array of factors including but not limited to skills, experience and actual hiring location. Your recruiter can share more information about the specific offer for the job location during the hiring process. Additional Information: Wolters Kluwer offers a wide variety of competitive benefits and programs to help meet your needs and balance your work and personal life, including but not limited to: Medical, Dental, & Vision Plans, 401(k), FSA/HSA, Commuter Benefits, Tuition Assistance Plan, Vacation and Sick Time, and Paid Parental Leave. Full details of our benefits are available upon request.

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board