Owns moderately complex components within Compute/Imaging platform services; leads team-level improvements to integration frameworks and developer tooling. Performs deep debugging across a bounded set of Operating System, Infrastructure Orchestration, distributed services, driving fixes that protect downstream consumers and upgrade paths. Analyzes usage, performance, and error budgets for specific platform surfaces; implements targeted resilience and capacity optimizations. Authors and curates team-scoped documentation, samples, and adoption guidance. Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability. Delivers features and load/performance tests; leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing; and reviews peers' implementations for scalability compliance. Builds fault-tolerant paths (redundancy, replication, automatic failover), applies recovery oriented principles, and implements retries, circuit breakers, and timeouts. Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry; authors runbooks and participates in incident response and RCAs. Implements standard replication and synchronization, develops automation/IaC for troubleshooting and maintenance, and applies advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met. Responsibilities Key Responsibilities Platform Software Development: Own a bounded Compute/Imaging platform component (service module area) and evolve its contracts for multi-tenant use. Perform deep debugging across a limited-service graph; drive compatibility-safe remediation plans. Implement targeted resilience/capacity patterns and document adoption guidance for team consumers. Software Development and Coding - Design, Testing, and Optimization: Designs software solutions and analyzes and helps identify requirements to achieve business and operational goals, independently. Adheres to and suggests improvements to all phases of the software development lifecycle. Utilizes working knowledge to develop new software features and enhancements following design specifications and develops documents to clarify software design and code. Leads code reviews in designated areas to help drive improvements. Conducts debugging and troubleshooting to identify and fix moderately complex software issues. Develops fixes for identified issues. Implements software testing (e.g., functional and non-functional testing), quality assurance processes, software error logging, monitoring, and observability for effective debugging, ensuring review by manager and/or lead throughout the process. Exercises judgment and discretion to conduct performance profiling and optimization of coding. Troubleshoots and resolves moderately complex issues related to application programming interface (API) functionality and integration. Implements moderately complex API versioning, lifecycle, and interoperability strategies. Software Architecture - Software System Structural Design: Designs and develops software, systems, and services aligned to pre-defined system architecture. Develops working knowledge of software architecture decisions and best practices. Collaborates with team leads to review work to ensure alignment with software architecture. Implements moderately complex performance optimization and scalability strategies in software design. Issue/Defect Collaboration - Software Products Support: Collaborates within and beyond immediate team to understand customer issues and align solutions. May provide technical guidance and support to customers regarding customer-reported issues, independently. Coaches, mentors, and guides others to advocate for customers' interests and suggests product enhancements based on feedback. Influences team to ensure customer satisfaction through timely resolution of issues and effective communication. Implements customer issue and/or defect handling and training processes, independently. Investigates and troubleshoots complex product maintenance issues to ensure customer agreement on short- and long-term solutions (e.g., future enhancements). Practices and Standards Compliance - Security and Compliance: Collaborates with the team to establish and follow development practices and coding standards. Exercises judgment and discretion to ensure code quality and adherence to broad acceptance criteria during development. Keeps up to date with industry best practices and applies them to software development processes. Implements secure coding practices to prevent security vulnerabilities. Development Operations - System Maintenance: Performs periodic maintenance and testing operations for systems that require upgrading or patching (e.g., for critical vulnerabilities). Exercises judgment and discretion to drive improvements, ensure automation, testing, and debugging of systems to ensure service/product availability, health, support, and reliability. Participate in Teams Oncall Rotations to provide reactive support for team's service area for critical incidents recovery. Core Responsibilities Planning & Execution: Independently manages work, monitoring timelines and deliverables to ensure projects or initiatives stay on track and meet requirements. Proactively prioritizes work and adapts to resource or timeline shifts, suggesting adjustments to maintain project efficiency. Collaboration & Partnership: Collaborates across teams to align on expectations and achieve shared objectives. Builds and maintains a comprehensive understanding of business, stakeholder, and/or customer needs to build and support effective partnerships. Actively listens to diverse perspectives and asks questions to ensure understanding of others. Problem Solving: Independently identifies and addresses standard and non-standard issues in accordance with standard practices, escalating more complex issues as appropriate. Analyzes data and/or information from multiple sources to troubleshoot standard and non-standard errors. Contributes to knowledge sharing and best practices. Continuous Learning: Embraces continuous learning by actively seeking to build knowledge and new skills and/or tools and staying current with industry trends and best practices. Seeks out and leverages feedback and training to improve skills. Contributes to a culture of continuous learning and knowledge sharing with team members. Continuous Improvement: Develops ideas and recommends updates to increase the efficiency and effectiveness of processes, protocols, and workflows within a team. Seeks input from team members on alternative approaches and methods for improving work. Requirements 5+ years' experience delivering and operating large scale, highly available distributed systems. Strong knowledge and interest in AI adoption including prompt engineering and agentic programming, with ChatGPT and Codex experience a plus. Strong knowledge of a base language such as Java, with a preference for functional programming language such as Scala. Strong knowledge of data structures, algorithms, operating systems, and distributed systems fundamentals. Experience with tools such as Terraform for Infrastructure as Code. Deep knowledge with networking protocols (TCP/IP, HTTP) and network architectures. Ability to design, troubleshoot and maintain networking infrastructure for high throughput use cases. Strong understanding of databases, storage, and distributed persistence technologies. Strong troubleshooting and performance tuning skills. Experience building multi-tenant, virtualized infrastructure a strong plus. Qualifications Disclaimer: Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements. Range and benefit information provided in this posting are specific to the stated locations only US: Hiring Range in USD from: $92,500 to $209,500 per annum. May be eligible for bonus and equity. Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following: 1. Medical, dental, and vision insurance, including expert medical opinion 2. Short term disability and long term disability 3. Life insurance and AD&D 4. Supplemental life insurance (Employee/Spouse/Child) 5. Health care and dependent care Flexible Spending Accounts 6. Pre-tax commuter and parking benefits 7. 401(k) Savings and Investment Plan with company match 8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week . click apply for full job details
09/24/2026
Full time
Owns moderately complex components within Compute/Imaging platform services; leads team-level improvements to integration frameworks and developer tooling. Performs deep debugging across a bounded set of Operating System, Infrastructure Orchestration, distributed services, driving fixes that protect downstream consumers and upgrade paths. Analyzes usage, performance, and error budgets for specific platform surfaces; implements targeted resilience and capacity optimizations. Authors and curates team-scoped documentation, samples, and adoption guidance. Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability. Delivers features and load/performance tests; leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing; and reviews peers' implementations for scalability compliance. Builds fault-tolerant paths (redundancy, replication, automatic failover), applies recovery oriented principles, and implements retries, circuit breakers, and timeouts. Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry; authors runbooks and participates in incident response and RCAs. Implements standard replication and synchronization, develops automation/IaC for troubleshooting and maintenance, and applies advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met. Responsibilities Key Responsibilities Platform Software Development: Own a bounded Compute/Imaging platform component (service module area) and evolve its contracts for multi-tenant use. Perform deep debugging across a limited-service graph; drive compatibility-safe remediation plans. Implement targeted resilience/capacity patterns and document adoption guidance for team consumers. Software Development and Coding - Design, Testing, and Optimization: Designs software solutions and analyzes and helps identify requirements to achieve business and operational goals, independently. Adheres to and suggests improvements to all phases of the software development lifecycle. Utilizes working knowledge to develop new software features and enhancements following design specifications and develops documents to clarify software design and code. Leads code reviews in designated areas to help drive improvements. Conducts debugging and troubleshooting to identify and fix moderately complex software issues. Develops fixes for identified issues. Implements software testing (e.g., functional and non-functional testing), quality assurance processes, software error logging, monitoring, and observability for effective debugging, ensuring review by manager and/or lead throughout the process. Exercises judgment and discretion to conduct performance profiling and optimization of coding. Troubleshoots and resolves moderately complex issues related to application programming interface (API) functionality and integration. Implements moderately complex API versioning, lifecycle, and interoperability strategies. Software Architecture - Software System Structural Design: Designs and develops software, systems, and services aligned to pre-defined system architecture. Develops working knowledge of software architecture decisions and best practices. Collaborates with team leads to review work to ensure alignment with software architecture. Implements moderately complex performance optimization and scalability strategies in software design. Issue/Defect Collaboration - Software Products Support: Collaborates within and beyond immediate team to understand customer issues and align solutions. May provide technical guidance and support to customers regarding customer-reported issues, independently. Coaches, mentors, and guides others to advocate for customers' interests and suggests product enhancements based on feedback. Influences team to ensure customer satisfaction through timely resolution of issues and effective communication. Implements customer issue and/or defect handling and training processes, independently. Investigates and troubleshoots complex product maintenance issues to ensure customer agreement on short- and long-term solutions (e.g., future enhancements). Practices and Standards Compliance - Security and Compliance: Collaborates with the team to establish and follow development practices and coding standards. Exercises judgment and discretion to ensure code quality and adherence to broad acceptance criteria during development. Keeps up to date with industry best practices and applies them to software development processes. Implements secure coding practices to prevent security vulnerabilities. Development Operations - System Maintenance: Performs periodic maintenance and testing operations for systems that require upgrading or patching (e.g., for critical vulnerabilities). Exercises judgment and discretion to drive improvements, ensure automation, testing, and debugging of systems to ensure service/product availability, health, support, and reliability. Participate in Teams Oncall Rotations to provide reactive support for team's service area for critical incidents recovery. Core Responsibilities Planning & Execution: Independently manages work, monitoring timelines and deliverables to ensure projects or initiatives stay on track and meet requirements. Proactively prioritizes work and adapts to resource or timeline shifts, suggesting adjustments to maintain project efficiency. Collaboration & Partnership: Collaborates across teams to align on expectations and achieve shared objectives. Builds and maintains a comprehensive understanding of business, stakeholder, and/or customer needs to build and support effective partnerships. Actively listens to diverse perspectives and asks questions to ensure understanding of others. Problem Solving: Independently identifies and addresses standard and non-standard issues in accordance with standard practices, escalating more complex issues as appropriate. Analyzes data and/or information from multiple sources to troubleshoot standard and non-standard errors. Contributes to knowledge sharing and best practices. Continuous Learning: Embraces continuous learning by actively seeking to build knowledge and new skills and/or tools and staying current with industry trends and best practices. Seeks out and leverages feedback and training to improve skills. Contributes to a culture of continuous learning and knowledge sharing with team members. Continuous Improvement: Develops ideas and recommends updates to increase the efficiency and effectiveness of processes, protocols, and workflows within a team. Seeks input from team members on alternative approaches and methods for improving work. Requirements 5+ years' experience delivering and operating large scale, highly available distributed systems. Strong knowledge and interest in AI adoption including prompt engineering and agentic programming, with ChatGPT and Codex experience a plus. Strong knowledge of a base language such as Java, with a preference for functional programming language such as Scala. Strong knowledge of data structures, algorithms, operating systems, and distributed systems fundamentals. Experience with tools such as Terraform for Infrastructure as Code. Deep knowledge with networking protocols (TCP/IP, HTTP) and network architectures. Ability to design, troubleshoot and maintain networking infrastructure for high throughput use cases. Strong understanding of databases, storage, and distributed persistence technologies. Strong troubleshooting and performance tuning skills. Experience building multi-tenant, virtualized infrastructure a strong plus. Qualifications Disclaimer: Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements. Range and benefit information provided in this posting are specific to the stated locations only US: Hiring Range in USD from: $92,500 to $209,500 per annum. May be eligible for bonus and equity. Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following: 1. Medical, dental, and vision insurance, including expert medical opinion 2. Short term disability and long term disability 3. Life insurance and AD&D 4. Supplemental life insurance (Employee/Spouse/Child) 5. Health care and dependent care Flexible Spending Accounts 6. Pre-tax commuter and parking benefits 7. 401(k) Savings and Investment Plan with company match 8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week . click apply for full job details
Req ID: 378685 NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Networking Advisor to join our team in Plano, Texas (US-TX), United States (US). This is a hybrid role, on-site at our client site several days per week. Only local candidates will be considered. The Disaster Recovery (DR) Specialist is responsible for leading the organization's disaster recovery and IT resilience program to ensure business continuity during system outages, cyber incidents, natural disasters, or other disruptive events. This role oversees the development, implementation, testing, and continuous improvement of disaster recovery strategies, recovery plans, and operational readiness across critical business and technology environments. The DR Manager collaborates with infrastructure, cybersecurity, cloud, application, network, and business stakeholders to establish recovery objectives, coordinate recovery exercises, manage crisis response activities, and ensure compliance with regulatory and organizational requirements. The role also provides leadership during disaster recovery events and drives continuous enhancement of recovery capabilities, governance, and resilience standards. The ideal candidate possesses strong expertise in disaster recovery planning, business continuity management, infrastructure operations, risk management, and stakeholder coordination, along with the ability to lead cross-functional teams in high-pressure situations. Basic Qualifications: 5+ years of Disaster Recovery/Risk Management Experience Additional Skills: Applied knowledge of risk management concepts Strong knowledge of systems and network administration (i.e., desktop, server) Knowledge and application of Globally Accepted Information Security Principles Strong knowledge of network security that pertains to communications, computer system environments and related infrastructures Thorough knowledge of server and desktop configurations that will protect systems from unauthorized access and software invasion Preferred: CISSP, GIAC, SSCP or CEH NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this role is $87,720 - $131,580. Actual compensation will depend on a number of factors, including the candidate's relevant experience, technical skills, and other qualifications. This position may also be eligible for incentive compensation based on individual and/or company performance. This position is eligible for company benefits including medical, dental, and vision insurance with an employer contribution, flexible spending or health savings account, life and AD&D insurance, short and long term disability coverage, paid time off, employee assistance, participation in a 401k program with company match, and additional voluntary or legally-required benefits. About NTT DATA NTT DATA is a $30 billion business and technology services leader, serving 75% of the Fortune Global 100. We are committed to accelerating client success and positively impacting society through responsible innovation. We are one of the world's leading AI and digital infrastructure providers, with unmatched capabilities in enterprise-scale AI, cloud, security, connectivity, data centers and application services. our consulting and Industry solutions help organizations and society move confidently and sustainably into the digital future. As a Global Top Employer, we have experts in more than 50 countries. We also offer clients access to a robust ecosystem of innovation centers as well as established and start-up partners. NTT DATA is a part of NTT Group, which invests over $3 billion each year in R&D. Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored to each client's needs. While many positions offer remote or hybrid work options, these arrangements are subject to change based on client requirements. For employees near an NTT DATA office or client site, in-office attendance may be required for meetings or events, depending on business needs. At NTT DATA, we are committed to staying flexible and meeting the evolving needs of both our clients and employees. NTT DATA recruiters will never ask for payment or banking information and will only email addresses. If you are requested to provide payment or disclose banking information, please submit a contact us form, NTT DATA endeavors to make accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact us at This contact information is for accommodation requests only and cannot be used to inquire about the status of applications. NTT DATA is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status. For our EEO Policy Statement, please click here. If you'd like more information on your EEO rights under the law, please click here. For Pay Transparency information, please click here.
09/24/2026
Full time
Req ID: 378685 NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Networking Advisor to join our team in Plano, Texas (US-TX), United States (US). This is a hybrid role, on-site at our client site several days per week. Only local candidates will be considered. The Disaster Recovery (DR) Specialist is responsible for leading the organization's disaster recovery and IT resilience program to ensure business continuity during system outages, cyber incidents, natural disasters, or other disruptive events. This role oversees the development, implementation, testing, and continuous improvement of disaster recovery strategies, recovery plans, and operational readiness across critical business and technology environments. The DR Manager collaborates with infrastructure, cybersecurity, cloud, application, network, and business stakeholders to establish recovery objectives, coordinate recovery exercises, manage crisis response activities, and ensure compliance with regulatory and organizational requirements. The role also provides leadership during disaster recovery events and drives continuous enhancement of recovery capabilities, governance, and resilience standards. The ideal candidate possesses strong expertise in disaster recovery planning, business continuity management, infrastructure operations, risk management, and stakeholder coordination, along with the ability to lead cross-functional teams in high-pressure situations. Basic Qualifications: 5+ years of Disaster Recovery/Risk Management Experience Additional Skills: Applied knowledge of risk management concepts Strong knowledge of systems and network administration (i.e., desktop, server) Knowledge and application of Globally Accepted Information Security Principles Strong knowledge of network security that pertains to communications, computer system environments and related infrastructures Thorough knowledge of server and desktop configurations that will protect systems from unauthorized access and software invasion Preferred: CISSP, GIAC, SSCP or CEH NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this role is $87,720 - $131,580. Actual compensation will depend on a number of factors, including the candidate's relevant experience, technical skills, and other qualifications. This position may also be eligible for incentive compensation based on individual and/or company performance. This position is eligible for company benefits including medical, dental, and vision insurance with an employer contribution, flexible spending or health savings account, life and AD&D insurance, short and long term disability coverage, paid time off, employee assistance, participation in a 401k program with company match, and additional voluntary or legally-required benefits. About NTT DATA NTT DATA is a $30 billion business and technology services leader, serving 75% of the Fortune Global 100. We are committed to accelerating client success and positively impacting society through responsible innovation. We are one of the world's leading AI and digital infrastructure providers, with unmatched capabilities in enterprise-scale AI, cloud, security, connectivity, data centers and application services. our consulting and Industry solutions help organizations and society move confidently and sustainably into the digital future. As a Global Top Employer, we have experts in more than 50 countries. We also offer clients access to a robust ecosystem of innovation centers as well as established and start-up partners. NTT DATA is a part of NTT Group, which invests over $3 billion each year in R&D. Whenever possible, we hire locally to NTT DATA offices or client sites. This ensures we can provide timely and effective support tailored to each client's needs. While many positions offer remote or hybrid work options, these arrangements are subject to change based on client requirements. For employees near an NTT DATA office or client site, in-office attendance may be required for meetings or events, depending on business needs. At NTT DATA, we are committed to staying flexible and meeting the evolving needs of both our clients and employees. NTT DATA recruiters will never ask for payment or banking information and will only email addresses. If you are requested to provide payment or disclose banking information, please submit a contact us form, NTT DATA endeavors to make accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact us at This contact information is for accommodation requests only and cannot be used to inquire about the status of applications. NTT DATA is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or protected veteran status. For our EEO Policy Statement, please click here. If you'd like more information on your EEO rights under the law, please click here. For Pay Transparency information, please click here.
Manages team delivering scalable distributed systems and components on a 2-4 quarter horizon. Standardizes engineering practices and scalability requirements across teams; oversees optimization for high throughput, hyper scale workloads; and ensures effective use of distributed state tools and data plane platforms. Guides teams to design fault tolerant, in service upgradable systems, set SLO aligned durability/availability targets, and implement resiliency mechanisms (load shedding, throttling, rate limiting). Provides oversight for KPIs, telemetry, and moderately complex dashboards; directs design of functional/correctness requirements, fault injection tests, and replication/synchronization strategies. Ensures proactive incident management, operational readiness, and on call coverage; drives encryption/access control practices, remediation plans, and compliance documentation. Oversees development and maintenance of automation/IaC and partners with teams on change management plans enabling safe patching, updates, and rollbacks. Responsibilities Key Responsibilities System Design & Architecture - System Scalability: Manages the development and implementation of scalable distributed systems and components across multiple teams, including the effective use of distributed state management tools. Oversees code and/or system optimization efforts for large-scale data processing and high-throughput requirements within and across teams to support hyper-scale systems. Guides teams to define scalability requirements for owned components and ensures design and implementation requirements are met. Manages the use of data plane platforms to effectively handle large-scale data retrieval, storage, and processing. Ensures team accurately designs performance and load testing. System Design & Architecture - System Reliability Design: Manages the strategy for building fault-tolerant components and systems capable of withstanding in-service updates by guiding the implementation of redundancy, replication, and automatic failover mechanisms. Develops design strategies for systems to effectively handle service disruptions (e.g., network partitions) by prioritizing consistency, availability, or partition tolerance. Leads implementation and optimization initiatives across teams for approaches to handle network unreliability, including load-shedding, throttling, and rate-limiting. Guides teams to design components and systems that are durable and adhere to service level objectives (SLOs), setting expectations for availability and durability of other computing services within the department. System Design & Architecture - System Reliability Performance: Provides oversight in defining key performance indicators (KPIs) and telemetry to identify gaps or issues in running systems. Oversees the building and customization of moderately complex dashboards, telemetry systems, and alerting mechanisms to proactively monitor components and system health. System Design & Architecture - Correctness / Availability: Oversees the design and implementation of functional and correctness requirements for feature sets and/or systems in new or existing systems. Guides teams to design complex test scenarios (e.g., fault-injection, brown-out) to evaluate system correctness. Directs implementation strategies for data replication and synchronization techniques to maintain data integrity and availability. Operational Troubleshooting & Incident Management: Guides teams to be proactive when diagnosing, debugging, and resolving issues in active components and systems to support ongoing operation. Ensures teams leverage expertise to prevent interruptions, ensuring no maintenance windows are required for customers and users when resolving issues. Oversees operational readiness protocol and ensures teams remain knowledgeable of owned components and systems to support effective troubleshooting and performance. Oversees and approves schedules for operational support rotations. Compliance & Security: Oversees implementation of robust security measures to protect data and applications in multi-tenant environments, ensuring team strategies incorporate encryption techniques and access controls. Directs execution of remediation plans to address identified security gaps, promoting continuous improvement of security measures. Ensures comprehensive documentation and cloud infrastructure compliance with industry standards and regulations. Automation & Change Management: Oversees the development and maintenance of automation scripts and tools (e.g., Infrastructure as Code (IaC to manage cloud infrastructure. Works with teams to create and adhere to change management plans for patching, updating, and rolling back applications, and guides development of components to allow for automation of these processes. Core Responsibilities Planning & Execution: Manages multiple medium- to large-scale projects or initiatives across teams, ensuring timelines, deliverables, and budgets (when applicable) are monitored and met. Provides direction to teams on project work, setting priorities, and aligning with business needs. Guides teams on adjusting plans to accommodate resource or timeline changes. Collaboration & Partnership: Drives cross-functional partnerships to align on expectations and shared objectives across multiple teams. Coaches team members to develop strategic relationships with business leaders, stakeholders, and external partners to foster collaboration and long-term success. Promotes inclusivity by actively seeking and listening to diverse perspectives, ensuring others feel heard and respected. Problem Solving: Provides direction to multiple teams on addressing complex operational and/or technical issues, as well as guidance on analyzing complex data and/or information to identify solutions. Reviews and provides insights into unresolved or critical issues, helping teams to identify potential solutions. Continuous Learning: Models engaging in continuous learning to deepen expertise and stay ahead of industry trends, integrating best practices into strategic planning. Leverages feedback to drive personal and team skill improvements. Identifies skill gaps across teams and empowers team members to pursue learning and knowledge-sharing opportunities that build their expertise in new areas, coaching them to apply learnings to advance the organization. Continuous Improvement: Drives teams to collaborate on, develop, and implement ideas to increase the efficiency and effectiveness of processes, protocols, and workflows within and across teams, providing oversight. Guides teams to adopt new ideas for alternative approaches and methods and encourages feedback for continued improvement. Performance and Development: Drives performance across teams by providing feedback and coaching in alignment with performance management processes, guidelines, and expectations. Discusses development goals with team members, shares opportunities to facilitate career development, and ensures individual goals are aligned with broader organizational goals. Develops and manages talent acquisition pipeline by leading candidate interviews, monitoring promotion eligibility, and/or orchestrating talent resources. Minimum Job Qualifications Education and/or Experience: •9 years of experience in software development OR •Bachelor's of Technology (B.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 5 years of experience in software development OR •Bachelor's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 5 years of experience in software development OR •Master's of Technology (M.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 3 years of experience in software development OR •Master's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 3 years of experience in software development OR •Doctorate in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 1 year of experience in software development. Job Skills: Same skills as prior level, plus: •Agile Methodologies: Demonstrated ability to use agile methodologies to drive continuous improvement and product delivery. •Automation: Demonstrated ability in or knowledge of automation, including designing, implementing . click apply for full job details
09/24/2026
Full time
Manages team delivering scalable distributed systems and components on a 2-4 quarter horizon. Standardizes engineering practices and scalability requirements across teams; oversees optimization for high throughput, hyper scale workloads; and ensures effective use of distributed state tools and data plane platforms. Guides teams to design fault tolerant, in service upgradable systems, set SLO aligned durability/availability targets, and implement resiliency mechanisms (load shedding, throttling, rate limiting). Provides oversight for KPIs, telemetry, and moderately complex dashboards; directs design of functional/correctness requirements, fault injection tests, and replication/synchronization strategies. Ensures proactive incident management, operational readiness, and on call coverage; drives encryption/access control practices, remediation plans, and compliance documentation. Oversees development and maintenance of automation/IaC and partners with teams on change management plans enabling safe patching, updates, and rollbacks. Responsibilities Key Responsibilities System Design & Architecture - System Scalability: Manages the development and implementation of scalable distributed systems and components across multiple teams, including the effective use of distributed state management tools. Oversees code and/or system optimization efforts for large-scale data processing and high-throughput requirements within and across teams to support hyper-scale systems. Guides teams to define scalability requirements for owned components and ensures design and implementation requirements are met. Manages the use of data plane platforms to effectively handle large-scale data retrieval, storage, and processing. Ensures team accurately designs performance and load testing. System Design & Architecture - System Reliability Design: Manages the strategy for building fault-tolerant components and systems capable of withstanding in-service updates by guiding the implementation of redundancy, replication, and automatic failover mechanisms. Develops design strategies for systems to effectively handle service disruptions (e.g., network partitions) by prioritizing consistency, availability, or partition tolerance. Leads implementation and optimization initiatives across teams for approaches to handle network unreliability, including load-shedding, throttling, and rate-limiting. Guides teams to design components and systems that are durable and adhere to service level objectives (SLOs), setting expectations for availability and durability of other computing services within the department. System Design & Architecture - System Reliability Performance: Provides oversight in defining key performance indicators (KPIs) and telemetry to identify gaps or issues in running systems. Oversees the building and customization of moderately complex dashboards, telemetry systems, and alerting mechanisms to proactively monitor components and system health. System Design & Architecture - Correctness / Availability: Oversees the design and implementation of functional and correctness requirements for feature sets and/or systems in new or existing systems. Guides teams to design complex test scenarios (e.g., fault-injection, brown-out) to evaluate system correctness. Directs implementation strategies for data replication and synchronization techniques to maintain data integrity and availability. Operational Troubleshooting & Incident Management: Guides teams to be proactive when diagnosing, debugging, and resolving issues in active components and systems to support ongoing operation. Ensures teams leverage expertise to prevent interruptions, ensuring no maintenance windows are required for customers and users when resolving issues. Oversees operational readiness protocol and ensures teams remain knowledgeable of owned components and systems to support effective troubleshooting and performance. Oversees and approves schedules for operational support rotations. Compliance & Security: Oversees implementation of robust security measures to protect data and applications in multi-tenant environments, ensuring team strategies incorporate encryption techniques and access controls. Directs execution of remediation plans to address identified security gaps, promoting continuous improvement of security measures. Ensures comprehensive documentation and cloud infrastructure compliance with industry standards and regulations. Automation & Change Management: Oversees the development and maintenance of automation scripts and tools (e.g., Infrastructure as Code (IaC to manage cloud infrastructure. Works with teams to create and adhere to change management plans for patching, updating, and rolling back applications, and guides development of components to allow for automation of these processes. Core Responsibilities Planning & Execution: Manages multiple medium- to large-scale projects or initiatives across teams, ensuring timelines, deliverables, and budgets (when applicable) are monitored and met. Provides direction to teams on project work, setting priorities, and aligning with business needs. Guides teams on adjusting plans to accommodate resource or timeline changes. Collaboration & Partnership: Drives cross-functional partnerships to align on expectations and shared objectives across multiple teams. Coaches team members to develop strategic relationships with business leaders, stakeholders, and external partners to foster collaboration and long-term success. Promotes inclusivity by actively seeking and listening to diverse perspectives, ensuring others feel heard and respected. Problem Solving: Provides direction to multiple teams on addressing complex operational and/or technical issues, as well as guidance on analyzing complex data and/or information to identify solutions. Reviews and provides insights into unresolved or critical issues, helping teams to identify potential solutions. Continuous Learning: Models engaging in continuous learning to deepen expertise and stay ahead of industry trends, integrating best practices into strategic planning. Leverages feedback to drive personal and team skill improvements. Identifies skill gaps across teams and empowers team members to pursue learning and knowledge-sharing opportunities that build their expertise in new areas, coaching them to apply learnings to advance the organization. Continuous Improvement: Drives teams to collaborate on, develop, and implement ideas to increase the efficiency and effectiveness of processes, protocols, and workflows within and across teams, providing oversight. Guides teams to adopt new ideas for alternative approaches and methods and encourages feedback for continued improvement. Performance and Development: Drives performance across teams by providing feedback and coaching in alignment with performance management processes, guidelines, and expectations. Discusses development goals with team members, shares opportunities to facilitate career development, and ensures individual goals are aligned with broader organizational goals. Develops and manages talent acquisition pipeline by leading candidate interviews, monitoring promotion eligibility, and/or orchestrating talent resources. Minimum Job Qualifications Education and/or Experience: •9 years of experience in software development OR •Bachelor's of Technology (B.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 5 years of experience in software development OR •Bachelor's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 5 years of experience in software development OR •Master's of Technology (M.Tech) Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 3 years of experience in software development OR •Master's Degree in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 3 years of experience in software development OR •Doctorate in Computer Science, Computer Engineering, Software Engineering, Electrical/Electronics Engineering, Computer Information Systems, Information Systems, Information Technology, Telecommunications, Mathematics, Physics, or related field AND 1 year of experience in software development. Job Skills: Same skills as prior level, plus: •Agile Methodologies: Demonstrated ability to use agile methodologies to drive continuous improvement and product delivery. •Automation: Demonstrated ability in or knowledge of automation, including designing, implementing . click apply for full job details
Job Title: Specialist, Cyber Infrastructure Systems Engineer -System Administration Job Code: 44433 Job Location: Colorado Springs, CO Job Schedule: 5/8 - Employees work 8 hours per day - 5 days a week Job Description: We are seeking a skilled and experienced Cyber Infrastructure Systems Engineer (CISE) to join our Colorado Springs, CO team. This role requires an experience system administrator who will be responsible for the administration and maintenance of Windows-based servers and workstations across a DoD cleared enterprise environment. Duties include managing patching activities, monitoring system performance, and implementing system hardening measures to strengthen security, ensure compliance with organizational standards, and maintain operational reliability. The position also supports ongoing troubleshooting, vulnerability remediation, and continuous improvement of system configurations and processes. Essential Functions: Administer, maintain, and support Windows-based servers and workstations within a cleared enterprise or mission-critical DoD environment. Manage the deployment of operating system patches, security updates, and hotfixes in accordance with approved maintenance schedules and security requirements. Implement and sustain system hardening measures using applicable DoD security standards, STIGs, and organizational cybersecurity policies. Perform vulnerability identification and remediation activities to support compliance with RMF controls, POA&M tracking, and audit readiness efforts. Monitor system performance, availability, and operational health to ensure reliability and continuity of mission systems. Troubleshoot and resolve hardware, software, operating system, and access-related issues while adhering to established security procedures. Maintain accurate system configurations, change records, and technical documentation to support configuration management and compliance inspections. Coordinate with cybersecurity, network, and program teams to support accreditation activities, incident response, and secure system operations in classified or controlled environments. Qualifications: Bachelor's Degree with a minimum of 4 years of relevant experience, or Graduate Degree with a minimum of 2 years of relevant experience, or in lieu of a degree, a minimum of 8 years of prior related experience. Secret Clearance DoD 8570.01-M IAT Level 2 Required Professional Certifications: CCNA- Security, Security+ CE, or higher certification' Preferred Additional Skills: Experience serving as a DoD COMSEC Manager or supporting COMSEC account management, including handling, safeguarding, inventorying, and controlling cryptographic equipment and keying material in accordance with applicable DoD policies and procedures. Experience administering and supporting Red Hat Enterprise Linux (RHEL) systems in a DoD, federal, or other regulated environment, including patching, system hardening, user administration, and troubleshooting Utilizing virtualization technologies such as KVM, VMware, vSphere/vCenter, and microservices/containers (e.g., OpenShift, Docker) to enhance system functionality and efficiency In compliance with pay transparency requirements, the salary range for this role in Colorado state is $80,500-$149,500. This is not a guarantee of compensation or salary, as final offer amount may vary based on factors including but not limited to experience and geographic location. L3Harris also offers a variety of benefits, including health and disability insurance, 401(k) match, flexible spending accounts, EAP, education assistance, parental leave, paid time off, and company-paid holidays. The specific programs and options available to an employee may vary depending on date of hire, schedule type, and the applicability of collective bargaining agreements. The application window for this position will close on 11/10/2026.
09/24/2026
Full time
Job Title: Specialist, Cyber Infrastructure Systems Engineer -System Administration Job Code: 44433 Job Location: Colorado Springs, CO Job Schedule: 5/8 - Employees work 8 hours per day - 5 days a week Job Description: We are seeking a skilled and experienced Cyber Infrastructure Systems Engineer (CISE) to join our Colorado Springs, CO team. This role requires an experience system administrator who will be responsible for the administration and maintenance of Windows-based servers and workstations across a DoD cleared enterprise environment. Duties include managing patching activities, monitoring system performance, and implementing system hardening measures to strengthen security, ensure compliance with organizational standards, and maintain operational reliability. The position also supports ongoing troubleshooting, vulnerability remediation, and continuous improvement of system configurations and processes. Essential Functions: Administer, maintain, and support Windows-based servers and workstations within a cleared enterprise or mission-critical DoD environment. Manage the deployment of operating system patches, security updates, and hotfixes in accordance with approved maintenance schedules and security requirements. Implement and sustain system hardening measures using applicable DoD security standards, STIGs, and organizational cybersecurity policies. Perform vulnerability identification and remediation activities to support compliance with RMF controls, POA&M tracking, and audit readiness efforts. Monitor system performance, availability, and operational health to ensure reliability and continuity of mission systems. Troubleshoot and resolve hardware, software, operating system, and access-related issues while adhering to established security procedures. Maintain accurate system configurations, change records, and technical documentation to support configuration management and compliance inspections. Coordinate with cybersecurity, network, and program teams to support accreditation activities, incident response, and secure system operations in classified or controlled environments. Qualifications: Bachelor's Degree with a minimum of 4 years of relevant experience, or Graduate Degree with a minimum of 2 years of relevant experience, or in lieu of a degree, a minimum of 8 years of prior related experience. Secret Clearance DoD 8570.01-M IAT Level 2 Required Professional Certifications: CCNA- Security, Security+ CE, or higher certification' Preferred Additional Skills: Experience serving as a DoD COMSEC Manager or supporting COMSEC account management, including handling, safeguarding, inventorying, and controlling cryptographic equipment and keying material in accordance with applicable DoD policies and procedures. Experience administering and supporting Red Hat Enterprise Linux (RHEL) systems in a DoD, federal, or other regulated environment, including patching, system hardening, user administration, and troubleshooting Utilizing virtualization technologies such as KVM, VMware, vSphere/vCenter, and microservices/containers (e.g., OpenShift, Docker) to enhance system functionality and efficiency In compliance with pay transparency requirements, the salary range for this role in Colorado state is $80,500-$149,500. This is not a guarantee of compensation or salary, as final offer amount may vary based on factors including but not limited to experience and geographic location. L3Harris also offers a variety of benefits, including health and disability insurance, 401(k) match, flexible spending accounts, EAP, education assistance, parental leave, paid time off, and company-paid holidays. The specific programs and options available to an employee may vary depending on date of hire, schedule type, and the applicability of collective bargaining agreements. The application window for this position will close on 11/10/2026.
About the Team DoorDash Labs is an independent team within DoorDash. We're hiring a Senior Operations Specialist to lead a team of operators to execute against autonomy goals. If you have a passion for applying robotics solutions to a service loved by millions of people, then we want to talk to you! About the Role We are looking for a driven Senior Operations Specialist who is excited about managing a dynamic operations team leading the forefront of delivery technology. You will manage a PM shift (3:30-12AM MST) of our operations team supporting autonomous testing, bringing autonomous technology from ideation to commercialization. You will report to a Specialist Operations Manager or Operations Manager on our robotics team in our DoorDash Labs organization. This job is 100% onsite at our Operations Hub in Mesa, AZ with some duties in Chandler and Tempe, AZ. You're excited about this opportunity because you will Manage a component of our operations Effectively communicate status updates to key stakeholders, including timely identification and escalation/resolution of issues Serve as an escalation path for issues in the field, to include incidents and onsite presence depending on severity Maintain an "and not either or" attitude when it comes to balancing business needs, costs, and safety Be an expert in our autonomous technology and help track and convey changes to it Strive to push metrics and add to them personally We're excited about you because you have 4+ years of experience (preferably in the autonomous industry or with strong interest) Valid U.S. drivers license required and a successful completion of MVR required Managed an operationally efficient team through multiple product phases A proven track record of operating in cross-functional initiatives and succeeding in a complex and fast moving environment An analytical mindset and can deliver actionable recommendations out of complex situations; Proficiency in gSuite, JIRA and SQL is preferred but not essential Notice to Applicants for Jobs Located in NYC or Remote Jobs Associated With Office in NYC Only We use Covey as part of our hiring and/or promotional process for jobs in NYC and certain features may qualify it as an AEDT in NYC. As part of the hiring and/or promotion process, we provide Covey with job requirements and candidate submitted applications. We began using Covey Scout for Inbound from August 21, 2023, through December 21, 2023, and resumed using Covey Scout for Inbound again on June 29, 2024. The Covey tool has been reviewed by an independent auditor. Results of the audit may be viewed here: Covey About DoorDash At DoorDash, our mission to empower local economies shapes how our team members move quickly, learn, and reiterate in order to make impactful decisions that display empathy for our range of users-from Dashers to merchant partners to consumers. We are a technology and logistics company that started by enabling door-to-door delivery, and we are looking for team members who can help us go from a company that is known as the place you order food to a company that people turn to for any and all goods. DoorDash is growing rapidly and changing constantly, which gives our team members the opportunity to share their unique perspectives, solve new challenges, and own their careers. We're committed to supporting employees' happiness, healthiness, and overall well-being by providing comprehensive benefits and perks including premium healthcare, wellness expense reimbursement, paid parental leave and more. Our Commitment to Diversity and Inclusion We're committed to growing and empowering a more inclusive community within our company, industry, and cities. That's why we hire and cultivate diverse teams of people from all backgrounds, experiences, and perspectives. We believe that true innovation happens when everyone has room at the table and the tools, resources, and opportunity to excel. Statement of Non-Discrimination: In keeping with our beliefs and goals, no employee or applicant will face discrimination or harassment based on: race, color, ancestry, national origin, religion, age, gender, marital/domestic partner status, sexual orientation, gender identity or expression, disability status, or veteran status. Above and beyond discrimination and harassment based on "protected categories," we also strive to prevent other subtler forms of inappropriate behavior (i.e., stereotyping) from ever gaining a foothold in our office. Whether blatant or hidden, barriers to success have no place at DoorDash. We value a diverse workforce - people who identify as women, non-binary or gender non-conforming, LGBTQIA+, American Indian or Native Alaskan, Black or African American, Hispanic or Latinx, Native Hawaiian or Other Pacific Islander, differently-abled, caretakers and parents, and veterans are strongly encouraged to apply. Thank you to the Level Playing Field Institute for this statement of non-discrimination. Pursuant to the San Francisco Fair Chance Ordinance, Los Angeles Fair Chance Initiative for Hiring Ordinance, and any other state or local hiring regulations, we will consider for employment any qualified applicant, including those with arrest and conviction records, in a manner consistent with the applicable regulation. If you need any accommodations, please inform your recruiting contact upon initial connection. Notice to Applicants for Jobs Located in NYC or Remote Jobs Associated With Office in NYC Only We used Covey as part of our hiring and/or promotional process for jobs in NYC and certain features may qualify it as an AEDT in NYC. As part of the hiring and/or promotion process, we provided Covey with job requirements and candidate submitted applications. We began using Covey Scout for Inbound from August 21, 2023, through December 21, 2023. We resumed using Covey Scout for Inbound again on June 29, 2024, and ceased using Covey Scout for Inbound on April 30, 2026. The Covey tool has been reviewed by an independent auditor. Results of the audit may be viewed here: Beware of recruitment scams: DoorDash, Deliveroo, and Wolt will never ask you to pay money or share sensitive financial information during hiring - learn more about our legitimate recruiting process at Recruitment Scam Awareness - DoorDash .
09/24/2026
Full time
About the Team DoorDash Labs is an independent team within DoorDash. We're hiring a Senior Operations Specialist to lead a team of operators to execute against autonomy goals. If you have a passion for applying robotics solutions to a service loved by millions of people, then we want to talk to you! About the Role We are looking for a driven Senior Operations Specialist who is excited about managing a dynamic operations team leading the forefront of delivery technology. You will manage a PM shift (3:30-12AM MST) of our operations team supporting autonomous testing, bringing autonomous technology from ideation to commercialization. You will report to a Specialist Operations Manager or Operations Manager on our robotics team in our DoorDash Labs organization. This job is 100% onsite at our Operations Hub in Mesa, AZ with some duties in Chandler and Tempe, AZ. You're excited about this opportunity because you will Manage a component of our operations Effectively communicate status updates to key stakeholders, including timely identification and escalation/resolution of issues Serve as an escalation path for issues in the field, to include incidents and onsite presence depending on severity Maintain an "and not either or" attitude when it comes to balancing business needs, costs, and safety Be an expert in our autonomous technology and help track and convey changes to it Strive to push metrics and add to them personally We're excited about you because you have 4+ years of experience (preferably in the autonomous industry or with strong interest) Valid U.S. drivers license required and a successful completion of MVR required Managed an operationally efficient team through multiple product phases A proven track record of operating in cross-functional initiatives and succeeding in a complex and fast moving environment An analytical mindset and can deliver actionable recommendations out of complex situations; Proficiency in gSuite, JIRA and SQL is preferred but not essential Notice to Applicants for Jobs Located in NYC or Remote Jobs Associated With Office in NYC Only We use Covey as part of our hiring and/or promotional process for jobs in NYC and certain features may qualify it as an AEDT in NYC. As part of the hiring and/or promotion process, we provide Covey with job requirements and candidate submitted applications. We began using Covey Scout for Inbound from August 21, 2023, through December 21, 2023, and resumed using Covey Scout for Inbound again on June 29, 2024. The Covey tool has been reviewed by an independent auditor. Results of the audit may be viewed here: Covey About DoorDash At DoorDash, our mission to empower local economies shapes how our team members move quickly, learn, and reiterate in order to make impactful decisions that display empathy for our range of users-from Dashers to merchant partners to consumers. We are a technology and logistics company that started by enabling door-to-door delivery, and we are looking for team members who can help us go from a company that is known as the place you order food to a company that people turn to for any and all goods. DoorDash is growing rapidly and changing constantly, which gives our team members the opportunity to share their unique perspectives, solve new challenges, and own their careers. We're committed to supporting employees' happiness, healthiness, and overall well-being by providing comprehensive benefits and perks including premium healthcare, wellness expense reimbursement, paid parental leave and more. Our Commitment to Diversity and Inclusion We're committed to growing and empowering a more inclusive community within our company, industry, and cities. That's why we hire and cultivate diverse teams of people from all backgrounds, experiences, and perspectives. We believe that true innovation happens when everyone has room at the table and the tools, resources, and opportunity to excel. Statement of Non-Discrimination: In keeping with our beliefs and goals, no employee or applicant will face discrimination or harassment based on: race, color, ancestry, national origin, religion, age, gender, marital/domestic partner status, sexual orientation, gender identity or expression, disability status, or veteran status. Above and beyond discrimination and harassment based on "protected categories," we also strive to prevent other subtler forms of inappropriate behavior (i.e., stereotyping) from ever gaining a foothold in our office. Whether blatant or hidden, barriers to success have no place at DoorDash. We value a diverse workforce - people who identify as women, non-binary or gender non-conforming, LGBTQIA+, American Indian or Native Alaskan, Black or African American, Hispanic or Latinx, Native Hawaiian or Other Pacific Islander, differently-abled, caretakers and parents, and veterans are strongly encouraged to apply. Thank you to the Level Playing Field Institute for this statement of non-discrimination. Pursuant to the San Francisco Fair Chance Ordinance, Los Angeles Fair Chance Initiative for Hiring Ordinance, and any other state or local hiring regulations, we will consider for employment any qualified applicant, including those with arrest and conviction records, in a manner consistent with the applicable regulation. If you need any accommodations, please inform your recruiting contact upon initial connection. Notice to Applicants for Jobs Located in NYC or Remote Jobs Associated With Office in NYC Only We used Covey as part of our hiring and/or promotional process for jobs in NYC and certain features may qualify it as an AEDT in NYC. As part of the hiring and/or promotion process, we provided Covey with job requirements and candidate submitted applications. We began using Covey Scout for Inbound from August 21, 2023, through December 21, 2023. We resumed using Covey Scout for Inbound again on June 29, 2024, and ceased using Covey Scout for Inbound on April 30, 2026. The Covey tool has been reviewed by an independent auditor. Results of the audit may be viewed here: Beware of recruitment scams: DoorDash, Deliveroo, and Wolt will never ask you to pay money or share sensitive financial information during hiring - learn more about our legitimate recruiting process at Recruitment Scam Awareness - DoorDash .
Job Description Summary For over forty years, HarbourVest has been home to a committed team of professionals with an entrepreneurial spirit and a desire to deliver impactful solutions to our clients and investing partners. As our global firm grows, we continue to add individuals who seek a collaborative, open-door culture that values diversity and innovative thinking. In our collegial environment that's marked by low turnover and high energy, you'll be inspired to grow and thrive. Here, you will be encouraged to build on your strengths and acquire new skills and experiences. We are committed to fostering an environment of inclusion that promotes mutual respect among all employees. Understanding and valuing these differences optimizes the potential of both the individual and the firm. HarbourVest is an equal opportunity employer. This position will be a hybrid work arrangement. You will receive 18 remote workdays per quarter to use at your discretion, subject to manager approval. For example, you may choose to work in the office 4 days per week and take one remote day weekly (typically 13 weeks per quarter), leaving 5 additional remote days to be used as needed. We are seeking a highly skilled Associate, Quantitative DevOps Engineer to join our Quantitative Investment Science organization. This role is ideal for a technology professional with a strong passion for data, automation, and modern DevOps practices who thrives in sophisticated, high-impact environments. The successful candidate will collaborate with quantitative researchers, data and application engineers, vendors, and IT partners. They will develop, deploy, and support robust CI/CD pipelines and cloud-based environments. You will play a critical role in supporting the full application and data lifecycle-from development through production-within our Azure-based Investment Data Analytics Platform. Your contributions will directly enhance development efficiency, system reliability, and operational excellence across the organization. The ideal candidate is someone who: Experience supporting technology applications or digital products using DevOps and process automation methods, preferably in data-focused or research-based environments. Demonstrated experience implementing and administering Git-based repositories, including branching and version control standards. Proven expertise in CI/CD pipeline development and deployment automation, including the use of approval gates. Strong scripting capabilities in Bash, Python, and/or PowerShell. Working knowledge of Snowflake and data platform ecosystems. Experience managing cloud infrastructure using infrastructure-as- code methodologies, ideally within Microsoft Azure. Proficiency with containerization technologies such as Docker or Podman. Strong background in configuration management, security, and operational rigor. Proven track record to support production environments, manage incidents, diagnose sophisticated issues, and coordinate resolution efforts. Demonstrated success delivering results on fast-paced, complicated projects. Excellent verbal and written communication skills, with the ability to engage effectively with technical and non-technical collaborators at all levels. What you will do: Build, implement, and maintain continuous integration and delivery workflows to support automated build, test, approval, and deployment processes across development, test, and production environments. Provide DevOps and process automation expertise to improve the efficiency, scalability, and reliability of data-centric and analytical applications. Support the ongoing maintenance, stability, and performance of a Microsoft Azure cloud. Administer and enforce Git-based source control, including branching strategies, versioning standards, and code promotion practices. Develop and maintain infrastructure-as-code (IaC) solutions for managing cloud storage, compute, and networking resources. Leverage containerization technologies (Docker or Podman) to standardize application deployments. Apply strong configuration management and security guidelines, ensuring alignment with enterprise standards. Provide production support, including issue triage, root cause analysis, prioritization, and coordination with analysts and engineers for timely resolution. Collaborate closely with data engineers, researchers, product, and platform engineering to support high-complexity, important initiatives. Contribute to the continuous improvement of DevOps standards, tools, and operational processes. What you bring: High energy and collaborative work style Proficient in using scripting languages to develop quick prototypes and proof-of-concepts A track record of completed high-intensity, high-complexity projects Experience with business intelligence tools such as Tableau, Power BI, or equivalent Skilled in implementing and managing DevOps practices, CI/CD pipelines Education Preferred: Bachelor of Arts (B.A) degree or equivalent experience Bachelor of Science (B.S) credential or equivalent experience Experience: 6 Years of DevOps experience in cloud based, software delivery Experience working within the financial services or investment management industry Base Salary Range $135,000.00 - $145,000.00 This USD base salary range represents only one component of total compensation for this role and is provided in accordance with local requirements. This role is eligible for a discretionary annual bonus, which is determined based on individual and overall firm performance. In addition to salary and bonus eligibility, total compensation may include participation in long-term reward programs (typically for more senior-level roles). We also offer a comprehensive total rewards package that may include retirement contributions, health insurance, income protection, paid time off and leave benefits as well as well-being programs. Our total rewards offerings are influenced by several business factors, and eligibility for certain components will vary by position and geography. Please note the posted ranges do not apply outside the U.S. and should not be converted to other currencies as a proxy for compensation in other countries.
09/24/2026
Full time
Job Description Summary For over forty years, HarbourVest has been home to a committed team of professionals with an entrepreneurial spirit and a desire to deliver impactful solutions to our clients and investing partners. As our global firm grows, we continue to add individuals who seek a collaborative, open-door culture that values diversity and innovative thinking. In our collegial environment that's marked by low turnover and high energy, you'll be inspired to grow and thrive. Here, you will be encouraged to build on your strengths and acquire new skills and experiences. We are committed to fostering an environment of inclusion that promotes mutual respect among all employees. Understanding and valuing these differences optimizes the potential of both the individual and the firm. HarbourVest is an equal opportunity employer. This position will be a hybrid work arrangement. You will receive 18 remote workdays per quarter to use at your discretion, subject to manager approval. For example, you may choose to work in the office 4 days per week and take one remote day weekly (typically 13 weeks per quarter), leaving 5 additional remote days to be used as needed. We are seeking a highly skilled Associate, Quantitative DevOps Engineer to join our Quantitative Investment Science organization. This role is ideal for a technology professional with a strong passion for data, automation, and modern DevOps practices who thrives in sophisticated, high-impact environments. The successful candidate will collaborate with quantitative researchers, data and application engineers, vendors, and IT partners. They will develop, deploy, and support robust CI/CD pipelines and cloud-based environments. You will play a critical role in supporting the full application and data lifecycle-from development through production-within our Azure-based Investment Data Analytics Platform. Your contributions will directly enhance development efficiency, system reliability, and operational excellence across the organization. The ideal candidate is someone who: Experience supporting technology applications or digital products using DevOps and process automation methods, preferably in data-focused or research-based environments. Demonstrated experience implementing and administering Git-based repositories, including branching and version control standards. Proven expertise in CI/CD pipeline development and deployment automation, including the use of approval gates. Strong scripting capabilities in Bash, Python, and/or PowerShell. Working knowledge of Snowflake and data platform ecosystems. Experience managing cloud infrastructure using infrastructure-as- code methodologies, ideally within Microsoft Azure. Proficiency with containerization technologies such as Docker or Podman. Strong background in configuration management, security, and operational rigor. Proven track record to support production environments, manage incidents, diagnose sophisticated issues, and coordinate resolution efforts. Demonstrated success delivering results on fast-paced, complicated projects. Excellent verbal and written communication skills, with the ability to engage effectively with technical and non-technical collaborators at all levels. What you will do: Build, implement, and maintain continuous integration and delivery workflows to support automated build, test, approval, and deployment processes across development, test, and production environments. Provide DevOps and process automation expertise to improve the efficiency, scalability, and reliability of data-centric and analytical applications. Support the ongoing maintenance, stability, and performance of a Microsoft Azure cloud. Administer and enforce Git-based source control, including branching strategies, versioning standards, and code promotion practices. Develop and maintain infrastructure-as-code (IaC) solutions for managing cloud storage, compute, and networking resources. Leverage containerization technologies (Docker or Podman) to standardize application deployments. Apply strong configuration management and security guidelines, ensuring alignment with enterprise standards. Provide production support, including issue triage, root cause analysis, prioritization, and coordination with analysts and engineers for timely resolution. Collaborate closely with data engineers, researchers, product, and platform engineering to support high-complexity, important initiatives. Contribute to the continuous improvement of DevOps standards, tools, and operational processes. What you bring: High energy and collaborative work style Proficient in using scripting languages to develop quick prototypes and proof-of-concepts A track record of completed high-intensity, high-complexity projects Experience with business intelligence tools such as Tableau, Power BI, or equivalent Skilled in implementing and managing DevOps practices, CI/CD pipelines Education Preferred: Bachelor of Arts (B.A) degree or equivalent experience Bachelor of Science (B.S) credential or equivalent experience Experience: 6 Years of DevOps experience in cloud based, software delivery Experience working within the financial services or investment management industry Base Salary Range $135,000.00 - $145,000.00 This USD base salary range represents only one component of total compensation for this role and is provided in accordance with local requirements. This role is eligible for a discretionary annual bonus, which is determined based on individual and overall firm performance. In addition to salary and bonus eligibility, total compensation may include participation in long-term reward programs (typically for more senior-level roles). We also offer a comprehensive total rewards package that may include retirement contributions, health insurance, income protection, paid time off and leave benefits as well as well-being programs. Our total rewards offerings are influenced by several business factors, and eligibility for certain components will vary by position and geography. Please note the posted ranges do not apply outside the U.S. and should not be converted to other currencies as a proxy for compensation in other countries.
Job Description Job Description PALO ALTO FIREWALL ENGINEER Position Description Description This is an opening for a Firewall Engineer/Team Lead to support a Department of State (DoS) Bureau of Diplomatic Technology (DT) program. This program provides transparent, interconnected systems and security supporting the DoS in successfully carrying out its U.S. foreign policy mission. DT provides enterprise architecture design, engineering, operations and maintenance support services for servers, networks, firewalls, and enterprise applications across the Department. Program is named "Vanguard" and is an IT consolidation consisting of the Department's servers, mainframes, network devices, network perimeter, anti-virus engineering, public key infrastructure (PKI)/biometrics/encryption, monitoring tools, telephony, mobile computing platform, virtual environment, and enclave design/security engineering. This is a firewall engineer position within the Vanguard 2025 program, providing general Tier II monitoring, configuration, and support to multiple firewalls and perimeter security systems. The position directly supports DoS on-site to provide perimeter security protection to over 80,000 customers globally. This is a hybrid position based out of Beltsville, MD with 3 days on-site. Shift is MIDs (11:00pm-7:30am), Sun-Thurs after training. During the 6-8 week training period, it will be 5 days on-site Mon-Fri (7:00am-3:30pm). Your responsibilities will include: Provides Tier 2 support in the monitoring, management, and troubleshooting of perimeter devices to include firewalls, proxies, and mail transport agents. Along with being a Team Lead for the MID shift. As the Lead you will be expectant to give a turnover of events for the past shift, create weekly Quad Chart, and be able to direct work to other firewall operations staff. Must be able to guide other staff on tasks assigned to the team. Review request for time off and stay in touch with the manager on all events happening on the shift. Monitor Service Now application for Firewall Operations incident and service requests. Implements firewall rules and policies, perform troubleshoots of firewall, email, and Proxy platforms for performance issues. Analyzes network traffic captures. Escalates issues as required to Tier 3 staff and monitors issues throughout a problem's life cycle. Performs recurring maintenance activities such as device reboots and software upgrades on perimeter devices. Records and reports on firewall operations, utilization, and maintenance. Collaborate across Bureaus and Agencies to implement and repair network changes as they relate to perimeter security devices. Support Diplomatic Security Computer Incident Response Team (CIRT) by implementing block requests for IP's, Websites, and Email addresses. Update architecture diagrams using Visio. Monitor and perform health checks on multiple perimeter security devices. Draft, coordinate and publish outage notifications as required. Update shift logs and provide reports to leadership daily. Maintain standard operating procedures (SOP), work instructions, and other working documents. Attend weekly teleconferences, onsite meetings, and participate in working groups as required. Qualifications Required Education & Experience BA degree and 6 years of experience; may accept additional experience in lieu of degree. In Depth experience using Palo Alto Firewalls and Panorama monitoring tools to troubleshoot and configure a Firewall infrastructure. Strong understanding of networking, proxy, and packet filtering technologies. Minimum 5 years IT Customer Support experience (Tier I and/or Tier II). First-hand experience with supporting the monitoring and configuration of Firewall/DMZ infrastructure including Network and Application Firewall Packet Filtering technologies (StoneGate, Palo Alto, FortiGate, Azure Firewall, Cloudflare, Cisco Email Security Appliances (ESA), A10 proxy devices). CLI experience preferred. Experienced in utilizing network monitoring tools such as Nagios and NeuralStar. Basic Microsoft Windows Server 2016 implementation and troubleshooting from a security/firewall perspective. Experienced with TCP/IP network implementation and troubleshooting. Required Clearance US Citizenship. An Interim Secret clearance is required to start work and requires eligibility to obtain a Top Secret clearance. Powered by JazzHR ZpLPHj5F0L
09/24/2026
Full time
Job Description Job Description PALO ALTO FIREWALL ENGINEER Position Description Description This is an opening for a Firewall Engineer/Team Lead to support a Department of State (DoS) Bureau of Diplomatic Technology (DT) program. This program provides transparent, interconnected systems and security supporting the DoS in successfully carrying out its U.S. foreign policy mission. DT provides enterprise architecture design, engineering, operations and maintenance support services for servers, networks, firewalls, and enterprise applications across the Department. Program is named "Vanguard" and is an IT consolidation consisting of the Department's servers, mainframes, network devices, network perimeter, anti-virus engineering, public key infrastructure (PKI)/biometrics/encryption, monitoring tools, telephony, mobile computing platform, virtual environment, and enclave design/security engineering. This is a firewall engineer position within the Vanguard 2025 program, providing general Tier II monitoring, configuration, and support to multiple firewalls and perimeter security systems. The position directly supports DoS on-site to provide perimeter security protection to over 80,000 customers globally. This is a hybrid position based out of Beltsville, MD with 3 days on-site. Shift is MIDs (11:00pm-7:30am), Sun-Thurs after training. During the 6-8 week training period, it will be 5 days on-site Mon-Fri (7:00am-3:30pm). Your responsibilities will include: Provides Tier 2 support in the monitoring, management, and troubleshooting of perimeter devices to include firewalls, proxies, and mail transport agents. Along with being a Team Lead for the MID shift. As the Lead you will be expectant to give a turnover of events for the past shift, create weekly Quad Chart, and be able to direct work to other firewall operations staff. Must be able to guide other staff on tasks assigned to the team. Review request for time off and stay in touch with the manager on all events happening on the shift. Monitor Service Now application for Firewall Operations incident and service requests. Implements firewall rules and policies, perform troubleshoots of firewall, email, and Proxy platforms for performance issues. Analyzes network traffic captures. Escalates issues as required to Tier 3 staff and monitors issues throughout a problem's life cycle. Performs recurring maintenance activities such as device reboots and software upgrades on perimeter devices. Records and reports on firewall operations, utilization, and maintenance. Collaborate across Bureaus and Agencies to implement and repair network changes as they relate to perimeter security devices. Support Diplomatic Security Computer Incident Response Team (CIRT) by implementing block requests for IP's, Websites, and Email addresses. Update architecture diagrams using Visio. Monitor and perform health checks on multiple perimeter security devices. Draft, coordinate and publish outage notifications as required. Update shift logs and provide reports to leadership daily. Maintain standard operating procedures (SOP), work instructions, and other working documents. Attend weekly teleconferences, onsite meetings, and participate in working groups as required. Qualifications Required Education & Experience BA degree and 6 years of experience; may accept additional experience in lieu of degree. In Depth experience using Palo Alto Firewalls and Panorama monitoring tools to troubleshoot and configure a Firewall infrastructure. Strong understanding of networking, proxy, and packet filtering technologies. Minimum 5 years IT Customer Support experience (Tier I and/or Tier II). First-hand experience with supporting the monitoring and configuration of Firewall/DMZ infrastructure including Network and Application Firewall Packet Filtering technologies (StoneGate, Palo Alto, FortiGate, Azure Firewall, Cloudflare, Cisco Email Security Appliances (ESA), A10 proxy devices). CLI experience preferred. Experienced in utilizing network monitoring tools such as Nagios and NeuralStar. Basic Microsoft Windows Server 2016 implementation and troubleshooting from a security/firewall perspective. Experienced with TCP/IP network implementation and troubleshooting. Required Clearance US Citizenship. An Interim Secret clearance is required to start work and requires eligibility to obtain a Top Secret clearance. Powered by JazzHR ZpLPHj5F0L
Job Description Job Description Cimarron is seeking a Global Network Operations and Security Center (GNOSC) Manager to support the Missile Defense Agency (MDA) on the Integrated Research and Development for Enterprise Solutions (IRES) contract at the Schriever Space Force Base in the Colorado Springs area. Key Duties: Direct all facets of the continuous 24x7x365 GNOSC environment. Provide administrative and technical oversight to Watch Officers and Mission Support Team (MST) Engineers. Serve as an escalation point for complex IT incidents, outages, and service degradations. Ensure that all incidents are effectively escalated, managed, and resolved within established service level agreements (SLAs). Provide full communication of status, remediation plans, and operational actions to executive leadership and the Government customer. Deliver post-event recommendations such as "Lessons Learned" and "Hot Wash" inputs. Oversee ITIL-based continual service improvement initiatives to minimize risk to services and operational networks. Coordinate the Mission Support Schedule with the MST Lead and Watch Officers to ensure all CIO assets directly support event requirements, exercises, and war games in a geographically distributed enterprise. Ensure team compliance with Tier III Information Assurance practices, IT security governance, and DoD 8570/8140 requirements. Champion metrics-based IT Operations and Maintenance (O&M) and oversee Quality Assurance/Quality Control Inspection processes. Required Skills, Experience, and Education: Due to facility security requirements, only U.S. citizens are eligible for consideration at the time. This position requires access to federal facilities. Candidates must possess a valid, unexpired Real ID-compliant driver's license or state-issued identification card at the time of hire. If you are unsure whether your ID is Real ID-compliant, please check for the star symbol in the upper portion of your driver's license or state ID. Active Secret Clearance. 12 or more years of general, full-time work experience (may be reduced with advanced education). 6 or more years of progressive experience in IT operations management, network control center (NCC), or enterprise operations center (EOC) roles. 2 or more years of direct personnel management or team leadership experience in a complex, 24/7 IT environment. Education or experience with ITIL framework and ITIL-based processes, to include continual service improvement, Incident, and change management. Current DoD 8570.01 IAT Level II Certification (e.g., CompTIA Security+ CE) or higher (IAM Level II/III preferred, such as CISM or CISSP). Desired Skills, Experience, and Education: Active DoW Top Secret Security Clearance. Bachelor's degree (or higher) in Computer Science, Information Technology, Cybersecurity, or a related field. Extensive experience leading metrics-based IT Operations and Maintenance (O&M) teams within a DoD or MDA environment. Advanced understanding of the ITIL framework with related certifications (e.g., ITIL 4 Foundation, ITIL 4 Managing Professional, or Strategic Leader). Familiar with enterprise monitoring, ticketing, and security toolsets utilized by the team, including Remedy, and SNMP monitoring tools (e.g. SolarWinds and Structure Ware Datacenter. Working technical background in network engineering, systems administration (Windows/Linux), or cybersecurity operations. Experience running operations centers for large complex organizations. Understanding of military organizations and structures Experience working with DISA and DREN.
09/24/2026
Full time
Job Description Job Description Cimarron is seeking a Global Network Operations and Security Center (GNOSC) Manager to support the Missile Defense Agency (MDA) on the Integrated Research and Development for Enterprise Solutions (IRES) contract at the Schriever Space Force Base in the Colorado Springs area. Key Duties: Direct all facets of the continuous 24x7x365 GNOSC environment. Provide administrative and technical oversight to Watch Officers and Mission Support Team (MST) Engineers. Serve as an escalation point for complex IT incidents, outages, and service degradations. Ensure that all incidents are effectively escalated, managed, and resolved within established service level agreements (SLAs). Provide full communication of status, remediation plans, and operational actions to executive leadership and the Government customer. Deliver post-event recommendations such as "Lessons Learned" and "Hot Wash" inputs. Oversee ITIL-based continual service improvement initiatives to minimize risk to services and operational networks. Coordinate the Mission Support Schedule with the MST Lead and Watch Officers to ensure all CIO assets directly support event requirements, exercises, and war games in a geographically distributed enterprise. Ensure team compliance with Tier III Information Assurance practices, IT security governance, and DoD 8570/8140 requirements. Champion metrics-based IT Operations and Maintenance (O&M) and oversee Quality Assurance/Quality Control Inspection processes. Required Skills, Experience, and Education: Due to facility security requirements, only U.S. citizens are eligible for consideration at the time. This position requires access to federal facilities. Candidates must possess a valid, unexpired Real ID-compliant driver's license or state-issued identification card at the time of hire. If you are unsure whether your ID is Real ID-compliant, please check for the star symbol in the upper portion of your driver's license or state ID. Active Secret Clearance. 12 or more years of general, full-time work experience (may be reduced with advanced education). 6 or more years of progressive experience in IT operations management, network control center (NCC), or enterprise operations center (EOC) roles. 2 or more years of direct personnel management or team leadership experience in a complex, 24/7 IT environment. Education or experience with ITIL framework and ITIL-based processes, to include continual service improvement, Incident, and change management. Current DoD 8570.01 IAT Level II Certification (e.g., CompTIA Security+ CE) or higher (IAM Level II/III preferred, such as CISM or CISSP). Desired Skills, Experience, and Education: Active DoW Top Secret Security Clearance. Bachelor's degree (or higher) in Computer Science, Information Technology, Cybersecurity, or a related field. Extensive experience leading metrics-based IT Operations and Maintenance (O&M) teams within a DoD or MDA environment. Advanced understanding of the ITIL framework with related certifications (e.g., ITIL 4 Foundation, ITIL 4 Managing Professional, or Strategic Leader). Familiar with enterprise monitoring, ticketing, and security toolsets utilized by the team, including Remedy, and SNMP monitoring tools (e.g. SolarWinds and Structure Ware Datacenter. Working technical background in network engineering, systems administration (Windows/Linux), or cybersecurity operations. Experience running operations centers for large complex organizations. Understanding of military organizations and structures Experience working with DISA and DREN.
Job Description Job Description About Us: NPO Torino is a specialized Digital Transformation company, recently acquired by the Fondo Italiano d'Investimento (Italian Investment Fund). With our headquarters in Turin and an operational presence in the United States and Brazil, we support companies in their technological evolution journey by combining consulting expertise with advanced IT solutions. We offer services in cloud computing, cybersecurity, AI for business, FinOps, system integration, and IT governance, featuring a customized and results-driven approach. Our mission is to simplify digital complexity and transform it into strategic value for enterprises. Backed by the Fondo Italiano d'Investimento , NPO Torino is accelerating its growth as a partner of choice for companies looking to tackle innovation challenges with vision, pragmatism, and sustainability. Salary range $80K to $92K. Work is partially remote with occasional travel to USA, and in Canada. Role Description: We are looking for a highly qualified Senior Network Security Engineer to join our Network & Security Business Unit. The professional will be responsible for the design, implementation, maintenance, and troubleshooting of complex network security infrastructures. The ideal candidate has deep, vertical expertise with leading security vendors (Fortinet, Palo Alto Networks, Cisco, F5) and a proven track record of managing modern architectures, including SASE solutions. Key Responsibilities: Design and manage security architectures based on Next-Generation Firewalls (NGFW). Implement and manage SASE (Secure Access Service Edge) solutions for distributed environments. Perform configuration , policy management, and advanced troubleshooting on multi-vendor equipment. Manage perimeter and internal security within Enterprise environments. Conduct log analysis, incident management, and system hardening. Provide technical mentoring for junior team members (optional, if required). Fundamental Technical Requirements (Must-Have) The candidate must demonstrate practical, in-depth, hands-on experience with the following technologies: Palo Alto Networks: Advanced configuration and management of NGFW (PA Series). Solid experience with Panorama for centralized management. Proven experience with SASE solutions (Prisma Access) and GlobalProtect. Fortinet: Advanced management of FortiGate (including VDOM, SD-WAN, SSL Inspection). Experience with the wider Fortinet ecosystem (FortiManager, FortiAnalyzer). Cisco Security: Experience with Cisco Security solutions (Firepower/FTD, ASA). Knowledge of Cisco ISE (Identity Services Engine) and TrustSec architectures is a strong plus. F5 Networks: Implementation and management of F5 BIG-IP solutions. Strong configuration and troubleshooting skills on LTM (Local Traffic Manager) and/or APM (Access Policy Manager) / AWAF modules. Soft Skills & Additional Requirements: Minimum of 5+ years of experience in Network Security Engineering roles. Excellent troubleshooting and problem-solving skills under pressure. Preferred Certifications (Nice-to-Have): Palo Alto: PCNSE (Palo Alto Networks Certified Network Security Engineer). Fortinet: NSE 4 / NSE 7. Cisco: CCNP Security or CCIE Security. F5: F5 Certified Administrator (F5-CA) or Technology Specialist (F5-CTS). Introduce Yourself (Screening Questions): "Do you have direct implementation experience with Prisma Access? Can you briefly describe a SASE architecture you have managed?" "Have you ever managed migrations between these vendors (e.g., from Cisco to Fortinet or vice versa)?" What We Offer: Salary commensurate with experience. Structured training and certification plans. Corporate benefits: Welfare Plan, Smart Working.
09/24/2026
Full time
Job Description Job Description About Us: NPO Torino is a specialized Digital Transformation company, recently acquired by the Fondo Italiano d'Investimento (Italian Investment Fund). With our headquarters in Turin and an operational presence in the United States and Brazil, we support companies in their technological evolution journey by combining consulting expertise with advanced IT solutions. We offer services in cloud computing, cybersecurity, AI for business, FinOps, system integration, and IT governance, featuring a customized and results-driven approach. Our mission is to simplify digital complexity and transform it into strategic value for enterprises. Backed by the Fondo Italiano d'Investimento , NPO Torino is accelerating its growth as a partner of choice for companies looking to tackle innovation challenges with vision, pragmatism, and sustainability. Salary range $80K to $92K. Work is partially remote with occasional travel to USA, and in Canada. Role Description: We are looking for a highly qualified Senior Network Security Engineer to join our Network & Security Business Unit. The professional will be responsible for the design, implementation, maintenance, and troubleshooting of complex network security infrastructures. The ideal candidate has deep, vertical expertise with leading security vendors (Fortinet, Palo Alto Networks, Cisco, F5) and a proven track record of managing modern architectures, including SASE solutions. Key Responsibilities: Design and manage security architectures based on Next-Generation Firewalls (NGFW). Implement and manage SASE (Secure Access Service Edge) solutions for distributed environments. Perform configuration , policy management, and advanced troubleshooting on multi-vendor equipment. Manage perimeter and internal security within Enterprise environments. Conduct log analysis, incident management, and system hardening. Provide technical mentoring for junior team members (optional, if required). Fundamental Technical Requirements (Must-Have) The candidate must demonstrate practical, in-depth, hands-on experience with the following technologies: Palo Alto Networks: Advanced configuration and management of NGFW (PA Series). Solid experience with Panorama for centralized management. Proven experience with SASE solutions (Prisma Access) and GlobalProtect. Fortinet: Advanced management of FortiGate (including VDOM, SD-WAN, SSL Inspection). Experience with the wider Fortinet ecosystem (FortiManager, FortiAnalyzer). Cisco Security: Experience with Cisco Security solutions (Firepower/FTD, ASA). Knowledge of Cisco ISE (Identity Services Engine) and TrustSec architectures is a strong plus. F5 Networks: Implementation and management of F5 BIG-IP solutions. Strong configuration and troubleshooting skills on LTM (Local Traffic Manager) and/or APM (Access Policy Manager) / AWAF modules. Soft Skills & Additional Requirements: Minimum of 5+ years of experience in Network Security Engineering roles. Excellent troubleshooting and problem-solving skills under pressure. Preferred Certifications (Nice-to-Have): Palo Alto: PCNSE (Palo Alto Networks Certified Network Security Engineer). Fortinet: NSE 4 / NSE 7. Cisco: CCNP Security or CCIE Security. F5: F5 Certified Administrator (F5-CA) or Technology Specialist (F5-CTS). Introduce Yourself (Screening Questions): "Do you have direct implementation experience with Prisma Access? Can you briefly describe a SASE architecture you have managed?" "Have you ever managed migrations between these vendors (e.g., from Cisco to Fortinet or vice versa)?" What We Offer: Salary commensurate with experience. Structured training and certification plans. Corporate benefits: Welfare Plan, Smart Working.
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
09/24/2026
Full time
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
Job Description Job Description We are seeking a Senior Network Security Engineer for an operations-first role supporting enterprise network security infrastructure across on-premises, remote-access, hybrid-cloud, and cloud-connected environments. This is not primarily an architecture/design role. The priority is a hands-on engineer who can administer, configure, maintain, troubleshoot, patch, upgrade, back up, validate, document, and operate production security platforms with minimal ramp-up. Firewall operations: hands-on Cisco and Palo Alto firewall administration, rule changes, NAT, troubleshooting, policy cleanup, upgrades, backups, logging, and production support. VPN / remote access: support for remote-access VPN, site-to-site VPN, user connectivity issues, certificates, authentication flows, and after-hours troubleshooting. RSA / MFA administration: RSA SecurID or equivalent MFA operations, token support, server administration, user troubleshooting, VPN integration, certificates, patching, backups, logs, and monitoring. Day-to-day operations: ticket resolution, monitoring alerts, health checks, change requests, incident support, maintenance windows, operational reporting, and customer support. Configuration and administration: installing, configuring, maintaining, patching, upgrading, backing up, validating, and troubleshooting assigned security platforms. Production troubleshooting: strong TCP/IP, DNS, routing, firewall logs, packet captures, VPN authentication, certificate, and connectivity troubleshooting. Documentation and process discipline: SOPs, runbooks, diagrams, change records, rollback plans, evidence collection, knowledge transfer, and formal change management. Federal/customer environment maturity: Public Trust eligibility, regulated-environment documentation, customer support, cross-team coordination, and comfort working with government stakeholders. The best candidate can credibly say: "I have operated enterprise Cisco and Palo Alto firewalls in production, handled firewall rule changes and troubleshooting, supported VPN users and site-to-site tunnels, administered or supported RSA/MFA tied to VPN access, followed formal change-management processes, maintained documentation and backups, and can step into daily operational support with minimal ramp-up." Scope and Role Boundaries Primary platforms include Cisco ASA/Firepower/FTD/FMC, Palo Alto NGFW/Panorama/GlobalProtect, remote-access and site-to-site VPN, RSA SecurID Authentication Manager or comparable MFA, monitoring/logging/SIEM integrations, and related network security controls. Coordinate with SOC/NOC, cloud, identity/directory, wireless/LAN, server, endpoint, system owner, application, governance, and vendor teams during changes, incidents, troubleshooting, compliance, and audit support. Cloudflare, Cisco ISE/NAC, secure web/email gateways, packet visibility tools, SD-WAN/SASE/ZTNA, AWS/Azure security, and F5/application-delivery awareness are useful where they intersect with assigned operational support, but the core need is firewall, VPN, RSA/MFA, and production operations. Key Responsibilities Provide daily, weekly, monthly, and annual operational support for assigned security systems, including tickets, alerts, health checks, email/phone support, metrics, status reporting, and operational validation. Administer and troubleshoot enterprise firewalls, including rule bases, NAT, segmentation, high availability, threat prevention, VPN integration, logging, secure baselines, rule reviews, recertification, cleanup, and decommissioning. Install, configure, maintain, patch, upgrade, back up, and validate firewall, VPN, MFA, and related network security systems in production environments. Support remote-access VPN, site-to-site VPN, partner connectivity, cloud connectivity, mobile/remote users, certificates, authentication policies, availability, utilization, and user access issues. Maintain and troubleshoot RSA SecurID Authentication Manager or equivalent MFA services, including servers/appliances, agents, certificates, HA, backups, logs, monitoring, directory integration, VPN authentication, and token lifecycle support. Respond to incidents, vulnerability notices, urgent requests, vendor advisories, PSIRT notices, system alerts, and emergency troubleshooting while minimizing service disruption. Use firewall logs, VPN logs, packet captures, SIEM data, monitoring tools, DNS/routing checks, and standard diagnostics to resolve complex connectivity, authentication, TLS/certificate, and application-flow issues. Create and maintain topology diagrams, equipment inventories, configurations, SOPs, runbooks, implementation plans, rollback plans, build/upgrade procedures, troubleshooting notes, and knowledge articles. Follow approved change, release, incident, problem, and configuration-management processes; prepare change records, peer-review materials, validation evidence, root-cause analysis, metrics, and audit artifacts. Support vulnerability remediation, POA&M tracking, continuous monitoring, compliance reviews, audit evidence collection, and coordination with ISSO, system owner, and security governance teams. Requirements 7+ years of experience in network security engineering, network infrastructure, cybersecurity infrastructure, or a closely related role. 5+ years of hands-on experience administering, maintaining, and troubleshooting enterprise firewall platforms in production environments. Hands-on experience with Cisco security technologies such as Cisco ASA, Firepower, FTD, FMC, AnyConnect/Secure Client, or equivalent Cisco firewall/VPN platforms. Hands-on experience with Palo Alto Networks technologies such as NGFW, Panorama, GlobalProtect, security profiles, App-ID/User-ID, logging, and policy optimization. Experience administering or supporting RSA SecurID Authentication Manager or comparable enterprise MFA/two-factor authentication platforms, including token support, server operations, patching/upgrades, backups, certificates, monitoring, and directory/VPN integration. Strong knowledge of firewall policy, NAT, VPNs, routing, DNS, DHCP, BGP, TLS/certificates, packet captures, log analysis, segmentation, high availability, and common network diagnostic tools. Experience with enterprise monitoring, logging, SIEM, alerting, vulnerability management, incident response, formal change management, and regulated-environment documentation. Ability to create clear technical documentation, support customers and stakeholders, prioritize operational work, communicate clearly, and coordinate across technical teams. Ability to obtain and maintain a Public Trust background investigation. Desired Certifications Relevant certifications are helpful but should not replace demonstrated hands-on experience. Examples include CCNP Security, CCIE Security, PCNSE, PCCSE, CISSP, CCSP, AWS Certified Security - Specialty, AWS Advanced Networking - Specialty, Microsoft Certified: Azure Security Engineer Associate, Microsoft Certified: Azure Network Engineer Associate, CompTIA Security+, CompTIA CySA+, GIAC certifications, or equivalent vendor/cloud certifications. Core Competencies Enterprise firewall engineering and policy lifecycle management VPN, remote access, RSA/MFA, and token lifecycle operations Cloudflare, edge security, secure access, and Zero Trust support Content filtering, secure web/email gateway, and NAC operations Hybrid-cloud network security and secure connectivity Monitoring, logging, SIEM integration, and incident response support Security visibility, packet analysis, and advanced troubleshooting Vulnerability remediation, compliance evidence, and POA&M support Change management, documentation, reporting, and operational metrics Technical leadership, customer support, and cross-team collaboration Benefits 401(k) 401(k) matching Dental insurance Flexible schedule Flexible spending account Health insurance Health savings account Life insurance Paid time off Professional development assistance Referral program Retirement plan Tuition reimbursement Vision insurance
09/24/2026
Full time
Job Description Job Description We are seeking a Senior Network Security Engineer for an operations-first role supporting enterprise network security infrastructure across on-premises, remote-access, hybrid-cloud, and cloud-connected environments. This is not primarily an architecture/design role. The priority is a hands-on engineer who can administer, configure, maintain, troubleshoot, patch, upgrade, back up, validate, document, and operate production security platforms with minimal ramp-up. Firewall operations: hands-on Cisco and Palo Alto firewall administration, rule changes, NAT, troubleshooting, policy cleanup, upgrades, backups, logging, and production support. VPN / remote access: support for remote-access VPN, site-to-site VPN, user connectivity issues, certificates, authentication flows, and after-hours troubleshooting. RSA / MFA administration: RSA SecurID or equivalent MFA operations, token support, server administration, user troubleshooting, VPN integration, certificates, patching, backups, logs, and monitoring. Day-to-day operations: ticket resolution, monitoring alerts, health checks, change requests, incident support, maintenance windows, operational reporting, and customer support. Configuration and administration: installing, configuring, maintaining, patching, upgrading, backing up, validating, and troubleshooting assigned security platforms. Production troubleshooting: strong TCP/IP, DNS, routing, firewall logs, packet captures, VPN authentication, certificate, and connectivity troubleshooting. Documentation and process discipline: SOPs, runbooks, diagrams, change records, rollback plans, evidence collection, knowledge transfer, and formal change management. Federal/customer environment maturity: Public Trust eligibility, regulated-environment documentation, customer support, cross-team coordination, and comfort working with government stakeholders. The best candidate can credibly say: "I have operated enterprise Cisco and Palo Alto firewalls in production, handled firewall rule changes and troubleshooting, supported VPN users and site-to-site tunnels, administered or supported RSA/MFA tied to VPN access, followed formal change-management processes, maintained documentation and backups, and can step into daily operational support with minimal ramp-up." Scope and Role Boundaries Primary platforms include Cisco ASA/Firepower/FTD/FMC, Palo Alto NGFW/Panorama/GlobalProtect, remote-access and site-to-site VPN, RSA SecurID Authentication Manager or comparable MFA, monitoring/logging/SIEM integrations, and related network security controls. Coordinate with SOC/NOC, cloud, identity/directory, wireless/LAN, server, endpoint, system owner, application, governance, and vendor teams during changes, incidents, troubleshooting, compliance, and audit support. Cloudflare, Cisco ISE/NAC, secure web/email gateways, packet visibility tools, SD-WAN/SASE/ZTNA, AWS/Azure security, and F5/application-delivery awareness are useful where they intersect with assigned operational support, but the core need is firewall, VPN, RSA/MFA, and production operations. Key Responsibilities Provide daily, weekly, monthly, and annual operational support for assigned security systems, including tickets, alerts, health checks, email/phone support, metrics, status reporting, and operational validation. Administer and troubleshoot enterprise firewalls, including rule bases, NAT, segmentation, high availability, threat prevention, VPN integration, logging, secure baselines, rule reviews, recertification, cleanup, and decommissioning. Install, configure, maintain, patch, upgrade, back up, and validate firewall, VPN, MFA, and related network security systems in production environments. Support remote-access VPN, site-to-site VPN, partner connectivity, cloud connectivity, mobile/remote users, certificates, authentication policies, availability, utilization, and user access issues. Maintain and troubleshoot RSA SecurID Authentication Manager or equivalent MFA services, including servers/appliances, agents, certificates, HA, backups, logs, monitoring, directory integration, VPN authentication, and token lifecycle support. Respond to incidents, vulnerability notices, urgent requests, vendor advisories, PSIRT notices, system alerts, and emergency troubleshooting while minimizing service disruption. Use firewall logs, VPN logs, packet captures, SIEM data, monitoring tools, DNS/routing checks, and standard diagnostics to resolve complex connectivity, authentication, TLS/certificate, and application-flow issues. Create and maintain topology diagrams, equipment inventories, configurations, SOPs, runbooks, implementation plans, rollback plans, build/upgrade procedures, troubleshooting notes, and knowledge articles. Follow approved change, release, incident, problem, and configuration-management processes; prepare change records, peer-review materials, validation evidence, root-cause analysis, metrics, and audit artifacts. Support vulnerability remediation, POA&M tracking, continuous monitoring, compliance reviews, audit evidence collection, and coordination with ISSO, system owner, and security governance teams. Requirements 7+ years of experience in network security engineering, network infrastructure, cybersecurity infrastructure, or a closely related role. 5+ years of hands-on experience administering, maintaining, and troubleshooting enterprise firewall platforms in production environments. Hands-on experience with Cisco security technologies such as Cisco ASA, Firepower, FTD, FMC, AnyConnect/Secure Client, or equivalent Cisco firewall/VPN platforms. Hands-on experience with Palo Alto Networks technologies such as NGFW, Panorama, GlobalProtect, security profiles, App-ID/User-ID, logging, and policy optimization. Experience administering or supporting RSA SecurID Authentication Manager or comparable enterprise MFA/two-factor authentication platforms, including token support, server operations, patching/upgrades, backups, certificates, monitoring, and directory/VPN integration. Strong knowledge of firewall policy, NAT, VPNs, routing, DNS, DHCP, BGP, TLS/certificates, packet captures, log analysis, segmentation, high availability, and common network diagnostic tools. Experience with enterprise monitoring, logging, SIEM, alerting, vulnerability management, incident response, formal change management, and regulated-environment documentation. Ability to create clear technical documentation, support customers and stakeholders, prioritize operational work, communicate clearly, and coordinate across technical teams. Ability to obtain and maintain a Public Trust background investigation. Desired Certifications Relevant certifications are helpful but should not replace demonstrated hands-on experience. Examples include CCNP Security, CCIE Security, PCNSE, PCCSE, CISSP, CCSP, AWS Certified Security - Specialty, AWS Advanced Networking - Specialty, Microsoft Certified: Azure Security Engineer Associate, Microsoft Certified: Azure Network Engineer Associate, CompTIA Security+, CompTIA CySA+, GIAC certifications, or equivalent vendor/cloud certifications. Core Competencies Enterprise firewall engineering and policy lifecycle management VPN, remote access, RSA/MFA, and token lifecycle operations Cloudflare, edge security, secure access, and Zero Trust support Content filtering, secure web/email gateway, and NAC operations Hybrid-cloud network security and secure connectivity Monitoring, logging, SIEM integration, and incident response support Security visibility, packet analysis, and advanced troubleshooting Vulnerability remediation, compliance evidence, and POA&M support Change management, documentation, reporting, and operational metrics Technical leadership, customer support, and cross-team collaboration Benefits 401(k) 401(k) matching Dental insurance Flexible schedule Flexible spending account Health insurance Health savings account Life insurance Paid time off Professional development assistance Referral program Retirement plan Tuition reimbursement Vision insurance
Charlie Health Engineering, Product & Design
New York, New York
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
09/24/2026
Full time
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
Job Description Job Description About the Company : Sungrow North America is a leading provider of renewable energy solutions, specializing in the development and manufacturing of photovoltaic inverters and energy storage systems. The company offers a comprehensive range of products and services designed to optimize the performance and efficiency of solar power installations. Sungrow North America aims to provide sustainable and reliable energy solutions to meet the growing demand for clean power and is known for its commitment to innovation, high-quality standards, and exceptional customer service. Security Engineer - Network & Identity: The Security Engineer (Network & Identity) is a hands-on engineering role within the IT team responsible for designing, implementing, securing, and automating Sungrow USA's network security, PKI and certificate management, and identity & access infrastructure across on-premises, cloud, and SaaS environments. This role serves as the technical owner for network security architecture, cryptographic services, certificate lifecycle management, authentication, and access controls. The position focuses on Zero Trust security, network segmentation, certificate-based authentication, and identity protection to reduce organizational risk and enable secure business operations and platform ownership rather than SOC operations, threat monitoring, or incident response. Essential Duties and Responsibilities: Network Security Design, implement, and maintain secure enterprise network architectures across corporate offices, data centers, cloud platforms, and remote workforce environments. Architect network segmentation, Zero Trust access controls, and secure connectivity standards using Fortinet and Zscaler security solutions. Develop and maintain Zero Trust architectures across network, identity, endpoint, application, and cloud environments. Design and administer secure remote access using Zscaler Private Access, VPN technologies, and identity-aware access controls. Manage firewall policies, network security controls, routing security, DNS security, and hybrid-cloud connectivity. Design and support Network Access Control architectures using IEEE 802.1X, RADIUS, and certificate-based authentication. Assess network security posture, develop remediation plans, and drive continuous security improvements. Cryptography, PKI & Certificate Management Own the enterprise PKI, cryptography, and certificate lifecycle management architecture, standards, and governance program. Design and manage certificate-based authentication and machine identity solutions for users, devices, servers, applications, cloud workloads, and network infrastructure across Azure, AWS, and hybrid environments. Implement and maintain certificate lifecycle automation using Microsoft Cloud PKI, Keyfactor, CyberArk Certificate Manager, EJBCA, DigiCert, AppViewX, or comparable platforms. Manage certificate issuance, enrollment, discovery, deployment, monitoring, renewal, revocation, auditing, and compliance across the enterprise. Design and support cryptographic services and certificate-based security controls, including TLS/mTLS, code signing, PKI trust hierarchies, certificate-based authentication, SCEP, PKCS, and machine identities. Establish PKI and cryptographic standards, key management practices, and security controls to support Zero Trust, regulatory compliance, and enterprise security requirements. Troubleshoot and resolve complex certificate, cryptographic, trust chain, authentication, and secure communications issues across enterprise systems and applications. Identity & Access Define authentication and authorization standards for workforce, partner, application, service, and machine identities. Design, implement, and maintain Microsoft Entra ID architecture, tenant governance, and identity security controls. Develop, test, and enforce Conditional Access policies and Zero Trust access controls. Implement and maintain MFA, passwordless authentication, phishing-resistant authentication, and Microsoft Entra ID Protection capabilities. Design and support enterprise SSO and federation integrations using SAML, OAuth 2.0, OpenID Connect, and SCIM. Implement least-privilege and risk-based access models across enterprise platforms. Administer RBAC, administrative separation, Microsoft Entra Privileged Identity Management, and least-privilege access controls. Govern application registrations, service principals, enterprise applications, API permissions, and managed identities. Support B2B collaboration, guest-user governance, external workforce access, and third-party identity integrations. Design and implement security controls across Microsoft Azure and AWS environments. Apply least privilege, RBAC, encryption, secrets management, and secure configuration standards to on-prem and cloud resources. Automate identity provisioning and deprovisioning, access governance, certificate management, configuration validation, and security operations. Create reusable secure-by-default templates and reduce manual administration through automation and orchestration Conduct access reviews, entitlement certifications, and identity governance activities. Education or Desired License and Certificates: Bachelor's degree in Computer Science, Information Technology, Cybersecurity, Engineering, or a related field, or equivalent professional experience. Microsoft Certified: Identity and Access Administrator Associate (SC-300) preferred. Microsoft Certified: Azure Security Engineer Associate (AZ-500) preferred. CCNA, Fortinet, Zscaler, AWS Security, CISSP, CISM, Terraform, or relevant PKI certification preferred Preferred Experience & Qualifications: 5+ years of experience in security engineering, identity & access management (IAM), network security, cloud security, or a related enterprise IT discipline. Hands-on experience with Microsoft Entra ID, including Conditional Access, MFA, SSO, Identity Protection, PIM, RBAC, identity governance, and modern authentication protocols (SAML, OAuth, OpenID Connect, SCIM). Experience designing, implementing, and securing enterprise identity, privileged access, and machine identity solutions across hybrid and multi-cloud environments. Hands-on experience with Fortinet, Zscaler (ZIA/ZPA), Zero Trust architectures, least-privilege access models, and network security controls. Experience designing and operating enterprise PKI, certificate lifecycle management, certificate-based authentication, and machine identity platforms such as Keyfactor, DigiCert, EJBCA, AppViewX, CyberArk Certificate Manager, or similar solutions. Experience securing Azure and AWS environments, including identity, networking, encryption, secrets management, logging, and security monitoring. Experience with PAM and IGA platforms such as CyberArk, Delinea, BeyondTrust, SailPoint, Saviynt, or similar technologies. Experience integrating identity, network, cloud, and security telemetry with SIEM and security operations platforms. Strong automation and Infrastructure as Code skills using PowerShell, Python, Microsoft Graph API, REST APIs, Terraform, or similar technologies. Strong troubleshooting skills across authentication, federation, certificates, PKI, network security, cloud access, application integrations, and enterprise identity services. Knowledge of cybersecurity and compliance frameworks including SOC 2, ISO/IEC 27001, NIST CSF, NIST 800-63, CIS Controls, Zero Trust, and NERC CIP. Competencies: Mandarin fluency preferred but not required. Strong analytical, troubleshooting, and problem-solving skills. Ability to work independently and collaboratively in a fast-paced environment. Excellent communication, stakeholder management, and technical documentation skills. Strong organization, attention to detail, initiative, and ownership. Ability to balance security, reliability, usability, scalability, and business requirements. Proactive approach to automation, standardization, and continuous improvement. Travel 5%-20% Work Location and Status: Full time, Hybrid at any Sungrow USA office in Phoenix, Costa Mesa, or Houston No visa sponsorship Compensation: Compensation commensurate with experience Competitive salary and annual bonus eligibility Comprehensive benefits package including health, dental, vision, and retirement plans Strong personal and company growth opportunities Sungrow is an equal opportunity employer. Due to strong interest in this position, Sungrow will only reach out to those candidates who best meet the requirements. Thank you for your interest in Sungrow.
09/24/2026
Full time
Job Description Job Description About the Company : Sungrow North America is a leading provider of renewable energy solutions, specializing in the development and manufacturing of photovoltaic inverters and energy storage systems. The company offers a comprehensive range of products and services designed to optimize the performance and efficiency of solar power installations. Sungrow North America aims to provide sustainable and reliable energy solutions to meet the growing demand for clean power and is known for its commitment to innovation, high-quality standards, and exceptional customer service. Security Engineer - Network & Identity: The Security Engineer (Network & Identity) is a hands-on engineering role within the IT team responsible for designing, implementing, securing, and automating Sungrow USA's network security, PKI and certificate management, and identity & access infrastructure across on-premises, cloud, and SaaS environments. This role serves as the technical owner for network security architecture, cryptographic services, certificate lifecycle management, authentication, and access controls. The position focuses on Zero Trust security, network segmentation, certificate-based authentication, and identity protection to reduce organizational risk and enable secure business operations and platform ownership rather than SOC operations, threat monitoring, or incident response. Essential Duties and Responsibilities: Network Security Design, implement, and maintain secure enterprise network architectures across corporate offices, data centers, cloud platforms, and remote workforce environments. Architect network segmentation, Zero Trust access controls, and secure connectivity standards using Fortinet and Zscaler security solutions. Develop and maintain Zero Trust architectures across network, identity, endpoint, application, and cloud environments. Design and administer secure remote access using Zscaler Private Access, VPN technologies, and identity-aware access controls. Manage firewall policies, network security controls, routing security, DNS security, and hybrid-cloud connectivity. Design and support Network Access Control architectures using IEEE 802.1X, RADIUS, and certificate-based authentication. Assess network security posture, develop remediation plans, and drive continuous security improvements. Cryptography, PKI & Certificate Management Own the enterprise PKI, cryptography, and certificate lifecycle management architecture, standards, and governance program. Design and manage certificate-based authentication and machine identity solutions for users, devices, servers, applications, cloud workloads, and network infrastructure across Azure, AWS, and hybrid environments. Implement and maintain certificate lifecycle automation using Microsoft Cloud PKI, Keyfactor, CyberArk Certificate Manager, EJBCA, DigiCert, AppViewX, or comparable platforms. Manage certificate issuance, enrollment, discovery, deployment, monitoring, renewal, revocation, auditing, and compliance across the enterprise. Design and support cryptographic services and certificate-based security controls, including TLS/mTLS, code signing, PKI trust hierarchies, certificate-based authentication, SCEP, PKCS, and machine identities. Establish PKI and cryptographic standards, key management practices, and security controls to support Zero Trust, regulatory compliance, and enterprise security requirements. Troubleshoot and resolve complex certificate, cryptographic, trust chain, authentication, and secure communications issues across enterprise systems and applications. Identity & Access Define authentication and authorization standards for workforce, partner, application, service, and machine identities. Design, implement, and maintain Microsoft Entra ID architecture, tenant governance, and identity security controls. Develop, test, and enforce Conditional Access policies and Zero Trust access controls. Implement and maintain MFA, passwordless authentication, phishing-resistant authentication, and Microsoft Entra ID Protection capabilities. Design and support enterprise SSO and federation integrations using SAML, OAuth 2.0, OpenID Connect, and SCIM. Implement least-privilege and risk-based access models across enterprise platforms. Administer RBAC, administrative separation, Microsoft Entra Privileged Identity Management, and least-privilege access controls. Govern application registrations, service principals, enterprise applications, API permissions, and managed identities. Support B2B collaboration, guest-user governance, external workforce access, and third-party identity integrations. Design and implement security controls across Microsoft Azure and AWS environments. Apply least privilege, RBAC, encryption, secrets management, and secure configuration standards to on-prem and cloud resources. Automate identity provisioning and deprovisioning, access governance, certificate management, configuration validation, and security operations. Create reusable secure-by-default templates and reduce manual administration through automation and orchestration Conduct access reviews, entitlement certifications, and identity governance activities. Education or Desired License and Certificates: Bachelor's degree in Computer Science, Information Technology, Cybersecurity, Engineering, or a related field, or equivalent professional experience. Microsoft Certified: Identity and Access Administrator Associate (SC-300) preferred. Microsoft Certified: Azure Security Engineer Associate (AZ-500) preferred. CCNA, Fortinet, Zscaler, AWS Security, CISSP, CISM, Terraform, or relevant PKI certification preferred Preferred Experience & Qualifications: 5+ years of experience in security engineering, identity & access management (IAM), network security, cloud security, or a related enterprise IT discipline. Hands-on experience with Microsoft Entra ID, including Conditional Access, MFA, SSO, Identity Protection, PIM, RBAC, identity governance, and modern authentication protocols (SAML, OAuth, OpenID Connect, SCIM). Experience designing, implementing, and securing enterprise identity, privileged access, and machine identity solutions across hybrid and multi-cloud environments. Hands-on experience with Fortinet, Zscaler (ZIA/ZPA), Zero Trust architectures, least-privilege access models, and network security controls. Experience designing and operating enterprise PKI, certificate lifecycle management, certificate-based authentication, and machine identity platforms such as Keyfactor, DigiCert, EJBCA, AppViewX, CyberArk Certificate Manager, or similar solutions. Experience securing Azure and AWS environments, including identity, networking, encryption, secrets management, logging, and security monitoring. Experience with PAM and IGA platforms such as CyberArk, Delinea, BeyondTrust, SailPoint, Saviynt, or similar technologies. Experience integrating identity, network, cloud, and security telemetry with SIEM and security operations platforms. Strong automation and Infrastructure as Code skills using PowerShell, Python, Microsoft Graph API, REST APIs, Terraform, or similar technologies. Strong troubleshooting skills across authentication, federation, certificates, PKI, network security, cloud access, application integrations, and enterprise identity services. Knowledge of cybersecurity and compliance frameworks including SOC 2, ISO/IEC 27001, NIST CSF, NIST 800-63, CIS Controls, Zero Trust, and NERC CIP. Competencies: Mandarin fluency preferred but not required. Strong analytical, troubleshooting, and problem-solving skills. Ability to work independently and collaboratively in a fast-paced environment. Excellent communication, stakeholder management, and technical documentation skills. Strong organization, attention to detail, initiative, and ownership. Ability to balance security, reliability, usability, scalability, and business requirements. Proactive approach to automation, standardization, and continuous improvement. Travel 5%-20% Work Location and Status: Full time, Hybrid at any Sungrow USA office in Phoenix, Costa Mesa, or Houston No visa sponsorship Compensation: Compensation commensurate with experience Competitive salary and annual bonus eligibility Comprehensive benefits package including health, dental, vision, and retirement plans Strong personal and company growth opportunities Sungrow is an equal opportunity employer. Due to strong interest in this position, Sungrow will only reach out to those candidates who best meet the requirements. Thank you for your interest in Sungrow.
LG Energy Solution Michigan, Inc.
Holland, Michigan
Job Description Job Description Title : Specialist I, Information Security Reports to: Security Manager Location: Holland, MI LG Energy Solution Michigan, Inc is a global leader in advanced lithium-ion battery technology for electric vehicle (EV) and energy storage applications. With growing operations across the United States, we are committed to innovation, quality, and powering the future of clean energy. Join a leader in advanced EV batteries. Summary: LG Energy Solution is seeking an experienced Information Security Specialist to support the design, implementation, and continuous improvement of physical security programs for LG Energy Solution, Michigan (Holland site).This role is responsible for safeguarding personnel, assets, and critical infrastructure across manufacturing plants and offices environments. The ideal candidate brings expertise in physical security systems, risk management, and enterprise security operations, along with the ability to collaborate across global teams and align with corporate security standards. Responsibilities: Support daily security operations across the facility, covering both physical and IT security domains Ensure alignment with global security policies and regulatory/compliance requirements (e.g., ISO 27001, KNCT law, internal standards) Collaborate with cross-functional teams including IT, HR, Legal, EHS, and external agencies Lead and support incident response and investigations across both physical and cyber domains Coordinate response efforts with internal stakeholders and external law enforcement when required Develop and maintain incident response procedures and ensure operational readiness Support ISO 27001 and other compliance initiatives across physical and IT security domains Assist in audit preparation, documentation, and evidence collection Develop, document, and maintain security standard operating procedures (SOPs) and guidelines Support user security awareness initiatives (e.g., phishing simulations, training programs) Conduct regular physical security risk assessments and site audits Oversee design, implementation, and maintenance of physical security systems, including: CCTV / Video Surveillance (including AI-enabled analytics) Access Control Systems (badge readers, vehicle gates, speed gates) Screening Systems (X-ray, metal detectors) Perimeter protection (fencing, barriers, intrusion detection sensors) Manage vendors and system integrators for installation, maintenance, and upgrades Oversee guard force operations, including staffing models, post orders, and SOP optimization Monitor and analyze security events using SIEM tools (e.g., Splunk) Assist in cybersecurity incident response, including initial triage, containment, and documentation Conduct vulnerability scans and support remediation efforts (e.g., Tenable Nessus) Track and report vulnerabilities and patch status; coordinate with IT teams for timely remediation Support administration and tuning of security tools (endpoint protection, DLP, NAC) Review and process firewall exception requests in accordance with security policies Implement and manage secure network infrastructure, including firewalls, VPNs, and segmentation Maintain cleanliness at work-site in accordance with 5S3R Standards: Sort, Set in order, Shine, Standardize, Sustain Right Location, Right Quantity, Right Container Follow LGESMI existing cleaning SOP's during downtime (e.g. line stop, waiting time, etc.), which includes cleaning your designated machine and/or surrounding work area Perform other duties as assigned Qualifications: Bachelor's degree in Information Security, Computer Science, Information Management Systems, Security Management or related field Experience in information security or IT, preferably in manufacturing or industrial environments is a plus. Experience supporting corporate, manufacturing, or industrial environment. Relevant certification: Security+, Network+ CySa+, CCNA, CISM, CISSP, CISA, etc. (Optional) Experience: 1 to 3+ years of relevant work experience in security, law enforcement Experience with physical security systems, including CCTV, access control, intrusion detection, and alarm systems. Experience conducting security assessments, incident investigations, or facility security operations is preferred. Skills: Strong knowledge of cybersecurity systems, Enterprise IT operation, network concepts & architecture, implementing and maintaining ISMS Ability to conduct risk assessments and interpret security data Ability to handle sensitive and confidential information Excellent problem-solving skills and adaptability to new systems and processes Strong written and verbal communication skills English (Fluent), Korean (Fluent, Optional) Benefits Overview: 100% employer-paid Medical, Dental, and Vision premiums for you and your family 100% employer-paid disability and life insurance Employer-supported childcare/babysitting programs Generous Paid Time Off / Holidays Opportunity to grow in a diverse work environment with a global company 401k Retirement savings and planning with a generous company match Equal Opportunity Employer LG Energy Solution is an Equal Opportunity Employer and considers all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected status under applicable law. Reasonable accommodations may be made for qualified individuals with disabilities.
09/24/2026
Full time
Job Description Job Description Title : Specialist I, Information Security Reports to: Security Manager Location: Holland, MI LG Energy Solution Michigan, Inc is a global leader in advanced lithium-ion battery technology for electric vehicle (EV) and energy storage applications. With growing operations across the United States, we are committed to innovation, quality, and powering the future of clean energy. Join a leader in advanced EV batteries. Summary: LG Energy Solution is seeking an experienced Information Security Specialist to support the design, implementation, and continuous improvement of physical security programs for LG Energy Solution, Michigan (Holland site).This role is responsible for safeguarding personnel, assets, and critical infrastructure across manufacturing plants and offices environments. The ideal candidate brings expertise in physical security systems, risk management, and enterprise security operations, along with the ability to collaborate across global teams and align with corporate security standards. Responsibilities: Support daily security operations across the facility, covering both physical and IT security domains Ensure alignment with global security policies and regulatory/compliance requirements (e.g., ISO 27001, KNCT law, internal standards) Collaborate with cross-functional teams including IT, HR, Legal, EHS, and external agencies Lead and support incident response and investigations across both physical and cyber domains Coordinate response efforts with internal stakeholders and external law enforcement when required Develop and maintain incident response procedures and ensure operational readiness Support ISO 27001 and other compliance initiatives across physical and IT security domains Assist in audit preparation, documentation, and evidence collection Develop, document, and maintain security standard operating procedures (SOPs) and guidelines Support user security awareness initiatives (e.g., phishing simulations, training programs) Conduct regular physical security risk assessments and site audits Oversee design, implementation, and maintenance of physical security systems, including: CCTV / Video Surveillance (including AI-enabled analytics) Access Control Systems (badge readers, vehicle gates, speed gates) Screening Systems (X-ray, metal detectors) Perimeter protection (fencing, barriers, intrusion detection sensors) Manage vendors and system integrators for installation, maintenance, and upgrades Oversee guard force operations, including staffing models, post orders, and SOP optimization Monitor and analyze security events using SIEM tools (e.g., Splunk) Assist in cybersecurity incident response, including initial triage, containment, and documentation Conduct vulnerability scans and support remediation efforts (e.g., Tenable Nessus) Track and report vulnerabilities and patch status; coordinate with IT teams for timely remediation Support administration and tuning of security tools (endpoint protection, DLP, NAC) Review and process firewall exception requests in accordance with security policies Implement and manage secure network infrastructure, including firewalls, VPNs, and segmentation Maintain cleanliness at work-site in accordance with 5S3R Standards: Sort, Set in order, Shine, Standardize, Sustain Right Location, Right Quantity, Right Container Follow LGESMI existing cleaning SOP's during downtime (e.g. line stop, waiting time, etc.), which includes cleaning your designated machine and/or surrounding work area Perform other duties as assigned Qualifications: Bachelor's degree in Information Security, Computer Science, Information Management Systems, Security Management or related field Experience in information security or IT, preferably in manufacturing or industrial environments is a plus. Experience supporting corporate, manufacturing, or industrial environment. Relevant certification: Security+, Network+ CySa+, CCNA, CISM, CISSP, CISA, etc. (Optional) Experience: 1 to 3+ years of relevant work experience in security, law enforcement Experience with physical security systems, including CCTV, access control, intrusion detection, and alarm systems. Experience conducting security assessments, incident investigations, or facility security operations is preferred. Skills: Strong knowledge of cybersecurity systems, Enterprise IT operation, network concepts & architecture, implementing and maintaining ISMS Ability to conduct risk assessments and interpret security data Ability to handle sensitive and confidential information Excellent problem-solving skills and adaptability to new systems and processes Strong written and verbal communication skills English (Fluent), Korean (Fluent, Optional) Benefits Overview: 100% employer-paid Medical, Dental, and Vision premiums for you and your family 100% employer-paid disability and life insurance Employer-supported childcare/babysitting programs Generous Paid Time Off / Holidays Opportunity to grow in a diverse work environment with a global company 401k Retirement savings and planning with a generous company match Equal Opportunity Employer LG Energy Solution is an Equal Opportunity Employer and considers all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected status under applicable law. Reasonable accommodations may be made for qualified individuals with disabilities.
Gravitee is a 2025 Gartner Magic Quadrant Leader , on a mission to govern the world's intelligence . We deliver the industry's most advanced platform for Any API, Any Event, and Any AI Agent , trusted by global leaders like Michelin, Roche, and Blue Yonder. Why join us? The Mission : We are the first to bridge traditional API Management with the new frontier of AI Agent Security The Momentum : A high-growth Leader - combining market credibility with startup speed The DNA : We hire people who Hold Nothing Back - passionate builders who want to redefine digital infrastructure Don't just watch the AI revolution. Build the infrastructure that controls and secures it. The Role We are looking for a Senior Software Engineer to build and maintain the identity and authorization features of Gravitee Access Management (AM) - across the AM runtime and the access-management experience in Gamma, Gravitee's next generation product surface. This is a new role. Today, AM engineering is based entirely in Europe. This hire establishes US-hours ownership of Level 3 and Level 4 authentication and authorization incidents, and adds delivery capacity toward AM parity in Gamma - part of building sustainable L3/L4 engineering capability in the US. You will split your time roughly 80% feature delivery and 20% L3/L4 support and bug fixing (it varies week to week), working as an embedded member of the AM team, which is based in Europe. What You Will Be Doing In this role, you will: Design and deliver features end to end, from discovery and technical design through implementation, testing, release, and iteration. Build and maintain identity and authorization features of Gravitee Access Management, across the AM runtime and the AM experience in Gamma. Implement and support OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0 as both IdP and SP, SCIM, and FAPI/CIBA/UMA profiles. Work with token and session semantics - JWT, JWKS, key rotation, revocation, introspection, MFA and step-up, WebAuthn/FIDO2, and IdP federation and social login. Keep security behavior and upgrades safe: standards compliance, secure defaults, certificate and secret handling, consent, audit logs, and defenses against token replay, SSRF, and account takeover. Own safe migrations and backward compatibility across MongoDB and JDBC, and support multi-domain, multi-region deployments and login/token endpoint performance. Own US-hours Level 3 and Level 4 escalations for AM customers as part of the L3 pager duty rotation. Use LLMs and AI-assisted development tools thoughtfully for prototyping, implementation, testing, debugging, and exploration, applying sound engineering judgment to validate AI-generated work. Write meaningful automated tests and contribute to reliable delivery practices. • Collaborate with product managers, designers, engineers, and technical leaders - including the AM team based in Europe - to discover effective solutions and improve them through code and design reviews. Share what you learn and help the team make practical choices as identity standards and protocols evolve. Essential Skills We are looking for evidence that you can succeed in the role, whether gained through employment, open-source work, or equivalent practical experience: 5+ years building and running production backend software, on a team that ships and supports its own product; you have personally resolved production incidents. Strong Java experience (C# accepted if the object-oriented depth is there), with Maven and a reactive stack such as Vert.x/RxJava. Deep working knowledge of identity standards: OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0, SCIM, and FAPI/CIBA/UMA profiles. • Solid grasp of token and session semantics: JWT, JWKS, rotation, revocation, introspection, MFA/step-up, WebAuthn/FIDO2, and IdP federation. A security-first mindset: secure defaults, certificate and secret handling, audit logging, and awareness of token replay, SSRF, and account-takeover risks. Experience with safe migrations and backward compatibility across persistent data stores such as MongoDB or JDBCbacked relational databases. Git-based workflow, code review, and writing your own automated tests. Hands-on experience using LLMs or AI coding assistants as part of an engineering workflow, combined with the judgment to review and improve their output. Clear communication, collaborative problem-solving, and the ability to take an ambiguous problem through to production. Desired Skills You do not need to match every item. We would be especially interested in experience with: • Experience at an API gateway, proxy, or service-mesh vendor, or on the API platform team of a large company (e.g., Kong, Google Apigee, MuleSoft, Tyk, Solo.io, Traefik, WSO2). Kubernetes operators and CRDs; OpenAPI tooling; service mesh or Envoy experience. Docker, Kubernetes, and cloud-native application delivery. Model Context Protocol (MCP), Agent2Agent (A2A), tool calling, LLM proxies, or other emerging AI protocols and standards. Prior production experience is not required. Building or operating LLM-powered applications, RAG systems, or agentic workflows - especially their security, governance, and observability needs. Open-source software or enterprise developer platforms. Who Thrives at Gravitee Our growth is powered by people who bring passion to what they build, professionalism to how they work, and a commitment to doing things well. You will thrive here if you: • Bring energy and a constructive attitude to the team. • Adapt quickly and enjoy learning unfamiliar technologies and domains. • Take ownership, communicate clearly, and follow through with urgency. • Balance delivery speed with thoughtful engineering judgment. • Start with the customer problem and care about the quality of the experience you create. • Enjoy working in a fast-moving, collaborative, international environment. Life at Gravitee At Gravitee, we invest in humans, not just roles. You'll get: • Salary of $160,000 • Competitive medical coverage. • Pension / 401(k) program options. • Stock options - you build it, you own it. • 25 days of holiday plus in-country national holidays. • Three mental health days and a wellness allowance. • Your birthday off. • A professional development budget to support your growth. • A hybrid work culture with hubs across regions. • Quarterly team events and an annual company offsite. • A collaborative, international company culture. • Opportunities to grow your scope and career as Gravitee grows. At Gravitee, we believe diverse perspectives make better products and stronger teams. No employee or applicant will be treated less favorably on the grounds of sex, marital status, race, color, nationality, ethnic or national origin, disability, gender, sexual orientation, gender identity, age, pregnancy or maternity, marital or civil partner status, religion, or belief. By applying, you consent to Gravitee storing and processing the personal information you submit as part of the recruitment process.
09/23/2026
Full time
Gravitee is a 2025 Gartner Magic Quadrant Leader , on a mission to govern the world's intelligence . We deliver the industry's most advanced platform for Any API, Any Event, and Any AI Agent , trusted by global leaders like Michelin, Roche, and Blue Yonder. Why join us? The Mission : We are the first to bridge traditional API Management with the new frontier of AI Agent Security The Momentum : A high-growth Leader - combining market credibility with startup speed The DNA : We hire people who Hold Nothing Back - passionate builders who want to redefine digital infrastructure Don't just watch the AI revolution. Build the infrastructure that controls and secures it. The Role We are looking for a Senior Software Engineer to build and maintain the identity and authorization features of Gravitee Access Management (AM) - across the AM runtime and the access-management experience in Gamma, Gravitee's next generation product surface. This is a new role. Today, AM engineering is based entirely in Europe. This hire establishes US-hours ownership of Level 3 and Level 4 authentication and authorization incidents, and adds delivery capacity toward AM parity in Gamma - part of building sustainable L3/L4 engineering capability in the US. You will split your time roughly 80% feature delivery and 20% L3/L4 support and bug fixing (it varies week to week), working as an embedded member of the AM team, which is based in Europe. What You Will Be Doing In this role, you will: Design and deliver features end to end, from discovery and technical design through implementation, testing, release, and iteration. Build and maintain identity and authorization features of Gravitee Access Management, across the AM runtime and the AM experience in Gamma. Implement and support OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0 as both IdP and SP, SCIM, and FAPI/CIBA/UMA profiles. Work with token and session semantics - JWT, JWKS, key rotation, revocation, introspection, MFA and step-up, WebAuthn/FIDO2, and IdP federation and social login. Keep security behavior and upgrades safe: standards compliance, secure defaults, certificate and secret handling, consent, audit logs, and defenses against token replay, SSRF, and account takeover. Own safe migrations and backward compatibility across MongoDB and JDBC, and support multi-domain, multi-region deployments and login/token endpoint performance. Own US-hours Level 3 and Level 4 escalations for AM customers as part of the L3 pager duty rotation. Use LLMs and AI-assisted development tools thoughtfully for prototyping, implementation, testing, debugging, and exploration, applying sound engineering judgment to validate AI-generated work. Write meaningful automated tests and contribute to reliable delivery practices. • Collaborate with product managers, designers, engineers, and technical leaders - including the AM team based in Europe - to discover effective solutions and improve them through code and design reviews. Share what you learn and help the team make practical choices as identity standards and protocols evolve. Essential Skills We are looking for evidence that you can succeed in the role, whether gained through employment, open-source work, or equivalent practical experience: 5+ years building and running production backend software, on a team that ships and supports its own product; you have personally resolved production incidents. Strong Java experience (C# accepted if the object-oriented depth is there), with Maven and a reactive stack such as Vert.x/RxJava. Deep working knowledge of identity standards: OAuth 2.0 and OIDC flows (authorization code + PKCE, client credentials, token exchange), SAML 2.0, SCIM, and FAPI/CIBA/UMA profiles. • Solid grasp of token and session semantics: JWT, JWKS, rotation, revocation, introspection, MFA/step-up, WebAuthn/FIDO2, and IdP federation. A security-first mindset: secure defaults, certificate and secret handling, audit logging, and awareness of token replay, SSRF, and account-takeover risks. Experience with safe migrations and backward compatibility across persistent data stores such as MongoDB or JDBCbacked relational databases. Git-based workflow, code review, and writing your own automated tests. Hands-on experience using LLMs or AI coding assistants as part of an engineering workflow, combined with the judgment to review and improve their output. Clear communication, collaborative problem-solving, and the ability to take an ambiguous problem through to production. Desired Skills You do not need to match every item. We would be especially interested in experience with: • Experience at an API gateway, proxy, or service-mesh vendor, or on the API platform team of a large company (e.g., Kong, Google Apigee, MuleSoft, Tyk, Solo.io, Traefik, WSO2). Kubernetes operators and CRDs; OpenAPI tooling; service mesh or Envoy experience. Docker, Kubernetes, and cloud-native application delivery. Model Context Protocol (MCP), Agent2Agent (A2A), tool calling, LLM proxies, or other emerging AI protocols and standards. Prior production experience is not required. Building or operating LLM-powered applications, RAG systems, or agentic workflows - especially their security, governance, and observability needs. Open-source software or enterprise developer platforms. Who Thrives at Gravitee Our growth is powered by people who bring passion to what they build, professionalism to how they work, and a commitment to doing things well. You will thrive here if you: • Bring energy and a constructive attitude to the team. • Adapt quickly and enjoy learning unfamiliar technologies and domains. • Take ownership, communicate clearly, and follow through with urgency. • Balance delivery speed with thoughtful engineering judgment. • Start with the customer problem and care about the quality of the experience you create. • Enjoy working in a fast-moving, collaborative, international environment. Life at Gravitee At Gravitee, we invest in humans, not just roles. You'll get: • Salary of $160,000 • Competitive medical coverage. • Pension / 401(k) program options. • Stock options - you build it, you own it. • 25 days of holiday plus in-country national holidays. • Three mental health days and a wellness allowance. • Your birthday off. • A professional development budget to support your growth. • A hybrid work culture with hubs across regions. • Quarterly team events and an annual company offsite. • A collaborative, international company culture. • Opportunities to grow your scope and career as Gravitee grows. At Gravitee, we believe diverse perspectives make better products and stronger teams. No employee or applicant will be treated less favorably on the grounds of sex, marital status, race, color, nationality, ethnic or national origin, disability, gender, sexual orientation, gender identity, age, pregnancy or maternity, marital or civil partner status, religion, or belief. By applying, you consent to Gravitee storing and processing the personal information you submit as part of the recruitment process.
Location: San Antonio Farinon Park, Dallas - 1201 Elm St At EY, we're all in to shape your future with confidence. We'll help you succeed in a globally connected powerhouse of diverse teams and take your career wherever you want it to go. Join EY and help to build a better working world. The opportunity We are offering you a demanding role. You will use the most advanced Quality models to implement the newest delivery excellence solutions. This position includes possibility to interact within international environment and work with the most recognizable and influential players in Delivery Excellence space. Your key responsibilities Responsible for designing, developing, implementing, and supporting enterprise API solutions using Apigee, ensuring secure, scalable, and high-performing API ecosystems. Accountable for API lifecycle management including API design, development, security, deployment, monitoring, versioning, and governance. Assess, design, build, test, deploy, and document API integrations between enterprise applications, cloud platforms, third-party vendors, and partner systems. Design and implement reusable API proxies, shared flows, security policies, and common integration components. Must be comfortable working in Agile methodology, driving technical discussions, gathering requirements from multiple stakeholders, and preparing technical solution specifications. Develop API management solutions including traffic management, authentication, authorization, rate limiting, caching, and threat protection. Identify root causes of production issues and provide technical solutions for performance optimization and operational stability. Forecast technical risks, dependencies, and mitigation plans and communicate them proactively with architects and delivery managers. Ensure API governance standards, security compliance, and enterprise integration best practices are followed across projects. Contribute to organization assets, accelerators, frameworks, reusable components, and process improvements. Deliver Proof of Concepts (PoCs), conduct code reviews, and mentor junior developers. Implement DevOps and CI/CD practices for API development, testing, deployment, and monitoring. Support integration testing, performance testing, security testing, and user acceptance testing activities. Collaborate with business, security, infrastructure, and application teams to deliver scalable API-led connectivity solutions. Key skills: Experience in API Management and Integration projects and hands-on experience in Apigee development. Must have implementation experience using Apigee in at least 2 enterprise-scale projects. Strong experience with Apigee Edge and/or Apigee X. Experience designing and implementing RESTful APIs and API-first architectures. Strong hands-on experience with API Proxy development, Shared Flows, Flow Hooks, API Products, Developers, Apps, and Monetization concepts. Experience implementing API security standards including OAuth 2.0, OpenID Connect, JWT, SAML, API Keys, mTLS, and SSL/TLS. Experience with API Gateway policies such as Spike Arrest, Quota, Caching, Threat Protection, Traffic Management, Message Transformation, and Logging. Strong knowledge of JavaScript and policy-based API development within Apigee. Experience integrating enterprise applications, ERP systems, CRM applications, SaaS platforms, and cloud-native services. Experience with REST, SOAP, JSON, XML, OpenAPI/Swagger specifications, Microservices, and Event-Driven Architecture. Hands-on experience troubleshooting API performance, latency, scalability, and security issues. Experience with logging and monitoring tools such as Splunk, ELK, Dynatrace, Grafana, Cloud Monitoring, or equivalent platforms. Knowledge of CI/CD tools including Jenkins, GitHub, GitLab, Azure DevOps, Maven, and automated deployment pipelines. Experience working with Kubernetes, Docker, and cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Understanding of API Governance, API Analytics, Developer Portals, and API Lifecycle Management. Multi-domain expertise and experience with other integration platforms is an added advantage. Strong written and verbal communication skills. Apigee API Engineer Certification or related Google Cloud certifications are preferred. Should be capable of mentoring team members and handling customer interactions independently. Skills and attributes for success: Strong communication skills and ability to collaborate with business delivery teams. Analytical thinking. Strong team-player and self-started attitude. Ability to create quality process documentation. Well organized, attentive to details attitude. To qualify for the role, you must have: College degree in Computer Science, Engineering, Information Technology, or related field. Minimum 3 years of experience in API development, integration, and API management projects. Strong understanding of API Security, API Governance, and Enterprise Integration Patterns. Experience handling production support, incident management, and root cause analysis. Ability to manage technical risks and customer escalations. Ability to interact effectively across business, technical, and leadership teams. Experience working in Agile Scrum and DevOps environments. Additionally, good to have: Experience with other API Management platforms such as IBM API Connect (APIC) and DataPower will be an added advantage. Experience with event streaming platforms such as Kafka or Pub/Sub. Knowledge of project management tools like JIRA, Azure DevOps, and Confluence. Exposure to DevOps automation, Infrastructure as Code (Terraform), and cloud-native architectures. Knowledge of SDLC, Waterfall, Agile, SAFe Agile, and Scrum methodologies. Experience with API monetization and partner onboarding solutions. Understanding of enterprise architecture frameworks and digital transformation programs. Ideally, you'll also have: Ability to define processes for teams. Experience in project audits and reviews. Ability to analyse risks and metrics of projects. What we look for: A Team of people with commercial acumen, technical experience and enthusiasm to learn new things in this fast-moving environment An opportunity to be a part of market-leading, multi-disciplinary team of 250+ professionals, in the only integrated global transaction business worldwide. Opportunities to work with EY ServiceNow practices globally with leading businesses across a range of industries What working at EY offers At EY, we're dedicated to helping our clients, from start-ups to Fortune 500 companies - and the work we do with them is as varied as they are. You get to work with inspiring and meaningful projects. Our focus is education and coaching alongside practical experience to ensure your personal development. We value our employees and you will be able to control your own development with an individual progression plan. You will quickly grow into a responsible role with challenging and stimulating assignments. Moreover, you will be part of an interdisciplinary environment that emphasizes high quality and knowledge exchange. Plus, we offer: Support, coaching and feedback from some of the most engaging colleagues around Opportunities to develop new skills and progress your career The freedom and flexibility to handle your role in a way that's right for you About EY As a global leader in assurance, tax, transaction and advisory services, we're using the finance products, expertise and systems we've developed to build a better working world. That starts with a culture that believes in giving you the training, opportunities and creative freedom to make things better. Whenever you join, however long you stay, the exceptional EY experience lasts a lifetime. And with a commitment to hiring and developing the most passionate people, we'll make our ambition to be the best employer by 2020 a reality. If you can confidently demonstrate that you meet the criteria above, please contact us as soon as possible. Join us in building a better working world. Apply now What we offer you At EY, we harness our collective strength to empower you to shape your future with confidence through professional growth, personal fulfillment and an inclusive culture. Learn more at We offer a comprehensive compensation and benefits package where you'll be rewarded based on your performance and recognized for the value you bring to the business. The base salary range for this job is: New York City, Boston, and Washington DC Metro Areas, Washington State, and Southern California offices - $80,300 to $149,100 Bay Area California offices - $83,700 to $155,300 All other offices locations in the US, including Sacramento - $67,000 to $136,800 Individual salaries within these ranges are determined through a wide variety of factors including but not limited to education, experience, knowledge, skills and geography. In addition, our Total Rewards package includes medical and dental coverage . click apply for full job details
09/23/2026
Full time
Location: San Antonio Farinon Park, Dallas - 1201 Elm St At EY, we're all in to shape your future with confidence. We'll help you succeed in a globally connected powerhouse of diverse teams and take your career wherever you want it to go. Join EY and help to build a better working world. The opportunity We are offering you a demanding role. You will use the most advanced Quality models to implement the newest delivery excellence solutions. This position includes possibility to interact within international environment and work with the most recognizable and influential players in Delivery Excellence space. Your key responsibilities Responsible for designing, developing, implementing, and supporting enterprise API solutions using Apigee, ensuring secure, scalable, and high-performing API ecosystems. Accountable for API lifecycle management including API design, development, security, deployment, monitoring, versioning, and governance. Assess, design, build, test, deploy, and document API integrations between enterprise applications, cloud platforms, third-party vendors, and partner systems. Design and implement reusable API proxies, shared flows, security policies, and common integration components. Must be comfortable working in Agile methodology, driving technical discussions, gathering requirements from multiple stakeholders, and preparing technical solution specifications. Develop API management solutions including traffic management, authentication, authorization, rate limiting, caching, and threat protection. Identify root causes of production issues and provide technical solutions for performance optimization and operational stability. Forecast technical risks, dependencies, and mitigation plans and communicate them proactively with architects and delivery managers. Ensure API governance standards, security compliance, and enterprise integration best practices are followed across projects. Contribute to organization assets, accelerators, frameworks, reusable components, and process improvements. Deliver Proof of Concepts (PoCs), conduct code reviews, and mentor junior developers. Implement DevOps and CI/CD practices for API development, testing, deployment, and monitoring. Support integration testing, performance testing, security testing, and user acceptance testing activities. Collaborate with business, security, infrastructure, and application teams to deliver scalable API-led connectivity solutions. Key skills: Experience in API Management and Integration projects and hands-on experience in Apigee development. Must have implementation experience using Apigee in at least 2 enterprise-scale projects. Strong experience with Apigee Edge and/or Apigee X. Experience designing and implementing RESTful APIs and API-first architectures. Strong hands-on experience with API Proxy development, Shared Flows, Flow Hooks, API Products, Developers, Apps, and Monetization concepts. Experience implementing API security standards including OAuth 2.0, OpenID Connect, JWT, SAML, API Keys, mTLS, and SSL/TLS. Experience with API Gateway policies such as Spike Arrest, Quota, Caching, Threat Protection, Traffic Management, Message Transformation, and Logging. Strong knowledge of JavaScript and policy-based API development within Apigee. Experience integrating enterprise applications, ERP systems, CRM applications, SaaS platforms, and cloud-native services. Experience with REST, SOAP, JSON, XML, OpenAPI/Swagger specifications, Microservices, and Event-Driven Architecture. Hands-on experience troubleshooting API performance, latency, scalability, and security issues. Experience with logging and monitoring tools such as Splunk, ELK, Dynatrace, Grafana, Cloud Monitoring, or equivalent platforms. Knowledge of CI/CD tools including Jenkins, GitHub, GitLab, Azure DevOps, Maven, and automated deployment pipelines. Experience working with Kubernetes, Docker, and cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Understanding of API Governance, API Analytics, Developer Portals, and API Lifecycle Management. Multi-domain expertise and experience with other integration platforms is an added advantage. Strong written and verbal communication skills. Apigee API Engineer Certification or related Google Cloud certifications are preferred. Should be capable of mentoring team members and handling customer interactions independently. Skills and attributes for success: Strong communication skills and ability to collaborate with business delivery teams. Analytical thinking. Strong team-player and self-started attitude. Ability to create quality process documentation. Well organized, attentive to details attitude. To qualify for the role, you must have: College degree in Computer Science, Engineering, Information Technology, or related field. Minimum 3 years of experience in API development, integration, and API management projects. Strong understanding of API Security, API Governance, and Enterprise Integration Patterns. Experience handling production support, incident management, and root cause analysis. Ability to manage technical risks and customer escalations. Ability to interact effectively across business, technical, and leadership teams. Experience working in Agile Scrum and DevOps environments. Additionally, good to have: Experience with other API Management platforms such as IBM API Connect (APIC) and DataPower will be an added advantage. Experience with event streaming platforms such as Kafka or Pub/Sub. Knowledge of project management tools like JIRA, Azure DevOps, and Confluence. Exposure to DevOps automation, Infrastructure as Code (Terraform), and cloud-native architectures. Knowledge of SDLC, Waterfall, Agile, SAFe Agile, and Scrum methodologies. Experience with API monetization and partner onboarding solutions. Understanding of enterprise architecture frameworks and digital transformation programs. Ideally, you'll also have: Ability to define processes for teams. Experience in project audits and reviews. Ability to analyse risks and metrics of projects. What we look for: A Team of people with commercial acumen, technical experience and enthusiasm to learn new things in this fast-moving environment An opportunity to be a part of market-leading, multi-disciplinary team of 250+ professionals, in the only integrated global transaction business worldwide. Opportunities to work with EY ServiceNow practices globally with leading businesses across a range of industries What working at EY offers At EY, we're dedicated to helping our clients, from start-ups to Fortune 500 companies - and the work we do with them is as varied as they are. You get to work with inspiring and meaningful projects. Our focus is education and coaching alongside practical experience to ensure your personal development. We value our employees and you will be able to control your own development with an individual progression plan. You will quickly grow into a responsible role with challenging and stimulating assignments. Moreover, you will be part of an interdisciplinary environment that emphasizes high quality and knowledge exchange. Plus, we offer: Support, coaching and feedback from some of the most engaging colleagues around Opportunities to develop new skills and progress your career The freedom and flexibility to handle your role in a way that's right for you About EY As a global leader in assurance, tax, transaction and advisory services, we're using the finance products, expertise and systems we've developed to build a better working world. That starts with a culture that believes in giving you the training, opportunities and creative freedom to make things better. Whenever you join, however long you stay, the exceptional EY experience lasts a lifetime. And with a commitment to hiring and developing the most passionate people, we'll make our ambition to be the best employer by 2020 a reality. If you can confidently demonstrate that you meet the criteria above, please contact us as soon as possible. Join us in building a better working world. Apply now What we offer you At EY, we harness our collective strength to empower you to shape your future with confidence through professional growth, personal fulfillment and an inclusive culture. Learn more at We offer a comprehensive compensation and benefits package where you'll be rewarded based on your performance and recognized for the value you bring to the business. The base salary range for this job is: New York City, Boston, and Washington DC Metro Areas, Washington State, and Southern California offices - $80,300 to $149,100 Bay Area California offices - $83,700 to $155,300 All other offices locations in the US, including Sacramento - $67,000 to $136,800 Individual salaries within these ranges are determined through a wide variety of factors including but not limited to education, experience, knowledge, skills and geography. In addition, our Total Rewards package includes medical and dental coverage . click apply for full job details
Description JOIN THE ECS4Kids TEAM ECS4Kids (formerly Episcopal Children's Services) is a 2026 Top Workplace that honors young children's immense potential by helping them enter school ready to learn. We cultivate lifelong learners by immersing children in enriching guided experiences and discovery-oriented approaches. Through programs like Voluntary Prekindergarten (VPK), Head Start, School Readiness, Child Care Resource & Referral (CCRR), and more, our dedicated professionals work with families and caregivers throughout Florida to promote children's well-being. These enable us to strengthen key areas including motor development, cognitive development, social and emotional development, language and communication skills, and problem-solving skills. Partner with us to help empower communities to rise above generational poverty with comprehensive early childhood education and holistic family support. We have career opportunities available in several counties throughout Northeast and Central Florida. ECS4Kids offers a competitive benefit package which includes: Medical, dental and vision insurance 403(b) plan with 5% employer match Employee Assistance Program (EAP) Long-term & short-term disability insurance Employer-paid life insurance Paid holidays Generous paid time off Career development Qualifying employer for Public Service Loan Forgiveness Program Salary commensurate with experience. GENERAL DESCRIPTION: The Vice President of Technology & Security serves as a strategic technology partner to executive leadership, translating agency goals into a multi-year technology and security roadmap. The Vice President provides strategic leadership and oversight of all IT and information security functions for Episcopal Children's Services (ECS4Kids), ensuring the reliability, security, compliance, and effectiveness of the agency's technology infrastructure, cloud platforms, telecommunications, and information assets across a 24-county service territory in Florida and Georgia. Serving as the agency's Information Systems Security Oficer (ISSO), the Vice President owns cybersecurity, data governance, business continuity/disaster recovery, vendor management, technology planning, and regulatory compliance, and protects sensitive child, family, employee, and agency information. The role develops technology policies, procedures, and security controls that align with organizational strategy, compliance requirements, and risk tolerance. The Vice President advises executive leadership and the Board of Directors on technology strategy, cybersecurity risk, and emerging technologies, and builds the technical and leadership capacity of a lean IT team supporting more than 1,000 end users across the agency's geographic footprint. In a department of this size, the Vice President is expected to operate as a player-coach - moving fluidly between strategic leadership and direct technical engagement as operational demands require. Major Responsibilities Partner with leadership to anticipate technology needs, develop a multi-year technology roadmap, and identify opportunities to improve business processes, service delivery, and efficiency through technology. Partner with Fiscal, Grants, and Program leadership to align technology investment with agency growth, funder priorities, and grant competitiveness. Serve as liaison between users, departments, vendors, and technical support providers, and report on technology performance, risk, and compliance to executive leadership and the Board. SaaS & Cloud Platform Management Oversee administration, configuration, security, and integration of the agency's SaaS ecosystem, including Microsoft 365, NetSuite (ERP), Paylocity (HCM), and other critical cloud platforms used across all program operations. Lead evaluation, selection, procurement, contract negotiation, and lifecycle management of SaaS and technology vendors, ensuring vendors meet security, compliance, and performance obligations. Maintain system access controls, identity management, and MFA enforcement across all cloud platforms; ensure proper ofboarding and access revocation protocols. Oversee implementation, upgrades, and vendor support of software, enterprise platforms, and network infrastructure, coordinating with the IT Manager and Specialists across the territory. Manage IT infrastructure - networks, cloud platforms, servers, telecommunications, and end-user devices - and oversee implementation, upgrades, and vendor support of software and enterprise platforms. Telecommunications & End User Support Manage IT asset lifecycle - end-user devices, mobile devices, peripherals - including procurement, deployment, tracking, and disposal, maintaining property custodian responsibilities. Oversee end user support operations, establishing service standards, escalation protocols, and performance metrics for the IT team's helpdesk and field support functions. Ensure the continuous availability of telecommunications infrastructure - internet connectivity, VoIP, and network services - across all 24 counties of the agency's service territory, minimizing downtime for more than 1,000 end users. Serve as ECS's Information Systems Security Oficer (ISSO), leading a comprehensive cybersecurity program: risk assessments, vulnerability and threat management, identity/access controls, MFA, endpoint protection, security monitoring, and awareness training (including for Microsoft 365, NetSuite, Paylocity, and other cloud systems). Lead incident response for cybersecurity events and data breaches, including investigation, containment, recovery, and required notiications. Provide direct technical support and hands-on problem-solving when escalated issues exceed team capacity or require executive-level intervention - including troubleshooting critical system outages, network failures, SaaS misconfigurations, or security incidents alongside IT staff. Serve as subject matter expert for third-party compliance documentation: dissect SOC 1 (Type II) and SOC 2 (Type II) reports from critical SaaS vendors, identify control deficiencies, evaluate system bridge letters, define required Complementary User Entity Controls (CUECs), and perform material weakness assessments to maintain internal compliance. Track and secure technology assets, mobile devices, and remote/cloud systems; maintain confidentiality of sensitive information encountered in the role. Establish agency-wide data governance (classification, privacy, retention, disposal, quality) and ensure database security and access controls. Ensure compliance with applicable state, federal, contractual, and funder technology and privacy requirements, including Head Start and Florida Early Learning standards; coordinate related audits, assessments, and cyber insurance requirements. Maintain documented IT policies, procedures, system/security documentation, and disaster recovery, incident response, and business continuity plans, including periodic testing and succession planning within IT. Direct technology procurement, vendor selection, contract negotiation, and lifecycle management, ensuring vendors meet security and compliance obligations. Develop and administer the annual IT operating and capital budget, including cost-beneit analysis and hardware planning; fulfill property custodian responsibilities. Supervise, coach, and develop IT staff, fostering a collaborative, customer-service-oriented, and continuously improving department. Evaluate emerging technologies (e.g., AI, automation) and establish governance and acceptable-use standards for their adoption. Serve in contract administrator/manager/supervisor and emergency duty roles as assigned; perform other related duties as assigned. These essential job functions are not a complete statement of duties. Employees may be required to perform other related duties as assigned. Requirements Education and Experience Bachelor's degree in Information Technology, Computer Science, Information Systems, Cybersecurity, Business Technology, or a related field. Seven (7) to ten (10) years of progressively responsible IT experience, including leadership over cybersecurity, SaaS/cloud platforms, telecommunications, end user support, vendor management, compliance, and strategic technology planning. Demonstrated experience reading and interpreting SOC 1 (Type II) and SOC 2 (Type II) audit reports, evaluating bridge letters, defining CUECs, and conducting material weakness assessments in a compliance context. Experience leading information security initiatives, technology projects, and business continuity/transformation efforts preferred. Experience in nonprofit, education, healthcare, human services, government, or similarly regulated environments preferred. (A comparable amount of training, education, or experience can be substituted for minimum qualifications.) Skills, Knowledge and Abilities Ability to provide strategic, forward-looking IT leadership, identifying risks and recommending proactive improvements. Demonstrated ability to operate effectively at both the strategic and operational level - equally comfortable presenting cybersecurity risk to the Board and troubleshooting a network outage or SaaS misconfiguration alongside IT staff. Ability to deine problems, interpret technical and diagrammatic instructions, and draw valid conclusions from data click apply for full job details
09/23/2026
Full time
Description JOIN THE ECS4Kids TEAM ECS4Kids (formerly Episcopal Children's Services) is a 2026 Top Workplace that honors young children's immense potential by helping them enter school ready to learn. We cultivate lifelong learners by immersing children in enriching guided experiences and discovery-oriented approaches. Through programs like Voluntary Prekindergarten (VPK), Head Start, School Readiness, Child Care Resource & Referral (CCRR), and more, our dedicated professionals work with families and caregivers throughout Florida to promote children's well-being. These enable us to strengthen key areas including motor development, cognitive development, social and emotional development, language and communication skills, and problem-solving skills. Partner with us to help empower communities to rise above generational poverty with comprehensive early childhood education and holistic family support. We have career opportunities available in several counties throughout Northeast and Central Florida. ECS4Kids offers a competitive benefit package which includes: Medical, dental and vision insurance 403(b) plan with 5% employer match Employee Assistance Program (EAP) Long-term & short-term disability insurance Employer-paid life insurance Paid holidays Generous paid time off Career development Qualifying employer for Public Service Loan Forgiveness Program Salary commensurate with experience. GENERAL DESCRIPTION: The Vice President of Technology & Security serves as a strategic technology partner to executive leadership, translating agency goals into a multi-year technology and security roadmap. The Vice President provides strategic leadership and oversight of all IT and information security functions for Episcopal Children's Services (ECS4Kids), ensuring the reliability, security, compliance, and effectiveness of the agency's technology infrastructure, cloud platforms, telecommunications, and information assets across a 24-county service territory in Florida and Georgia. Serving as the agency's Information Systems Security Oficer (ISSO), the Vice President owns cybersecurity, data governance, business continuity/disaster recovery, vendor management, technology planning, and regulatory compliance, and protects sensitive child, family, employee, and agency information. The role develops technology policies, procedures, and security controls that align with organizational strategy, compliance requirements, and risk tolerance. The Vice President advises executive leadership and the Board of Directors on technology strategy, cybersecurity risk, and emerging technologies, and builds the technical and leadership capacity of a lean IT team supporting more than 1,000 end users across the agency's geographic footprint. In a department of this size, the Vice President is expected to operate as a player-coach - moving fluidly between strategic leadership and direct technical engagement as operational demands require. Major Responsibilities Partner with leadership to anticipate technology needs, develop a multi-year technology roadmap, and identify opportunities to improve business processes, service delivery, and efficiency through technology. Partner with Fiscal, Grants, and Program leadership to align technology investment with agency growth, funder priorities, and grant competitiveness. Serve as liaison between users, departments, vendors, and technical support providers, and report on technology performance, risk, and compliance to executive leadership and the Board. SaaS & Cloud Platform Management Oversee administration, configuration, security, and integration of the agency's SaaS ecosystem, including Microsoft 365, NetSuite (ERP), Paylocity (HCM), and other critical cloud platforms used across all program operations. Lead evaluation, selection, procurement, contract negotiation, and lifecycle management of SaaS and technology vendors, ensuring vendors meet security, compliance, and performance obligations. Maintain system access controls, identity management, and MFA enforcement across all cloud platforms; ensure proper ofboarding and access revocation protocols. Oversee implementation, upgrades, and vendor support of software, enterprise platforms, and network infrastructure, coordinating with the IT Manager and Specialists across the territory. Manage IT infrastructure - networks, cloud platforms, servers, telecommunications, and end-user devices - and oversee implementation, upgrades, and vendor support of software and enterprise platforms. Telecommunications & End User Support Manage IT asset lifecycle - end-user devices, mobile devices, peripherals - including procurement, deployment, tracking, and disposal, maintaining property custodian responsibilities. Oversee end user support operations, establishing service standards, escalation protocols, and performance metrics for the IT team's helpdesk and field support functions. Ensure the continuous availability of telecommunications infrastructure - internet connectivity, VoIP, and network services - across all 24 counties of the agency's service territory, minimizing downtime for more than 1,000 end users. Serve as ECS's Information Systems Security Oficer (ISSO), leading a comprehensive cybersecurity program: risk assessments, vulnerability and threat management, identity/access controls, MFA, endpoint protection, security monitoring, and awareness training (including for Microsoft 365, NetSuite, Paylocity, and other cloud systems). Lead incident response for cybersecurity events and data breaches, including investigation, containment, recovery, and required notiications. Provide direct technical support and hands-on problem-solving when escalated issues exceed team capacity or require executive-level intervention - including troubleshooting critical system outages, network failures, SaaS misconfigurations, or security incidents alongside IT staff. Serve as subject matter expert for third-party compliance documentation: dissect SOC 1 (Type II) and SOC 2 (Type II) reports from critical SaaS vendors, identify control deficiencies, evaluate system bridge letters, define required Complementary User Entity Controls (CUECs), and perform material weakness assessments to maintain internal compliance. Track and secure technology assets, mobile devices, and remote/cloud systems; maintain confidentiality of sensitive information encountered in the role. Establish agency-wide data governance (classification, privacy, retention, disposal, quality) and ensure database security and access controls. Ensure compliance with applicable state, federal, contractual, and funder technology and privacy requirements, including Head Start and Florida Early Learning standards; coordinate related audits, assessments, and cyber insurance requirements. Maintain documented IT policies, procedures, system/security documentation, and disaster recovery, incident response, and business continuity plans, including periodic testing and succession planning within IT. Direct technology procurement, vendor selection, contract negotiation, and lifecycle management, ensuring vendors meet security and compliance obligations. Develop and administer the annual IT operating and capital budget, including cost-beneit analysis and hardware planning; fulfill property custodian responsibilities. Supervise, coach, and develop IT staff, fostering a collaborative, customer-service-oriented, and continuously improving department. Evaluate emerging technologies (e.g., AI, automation) and establish governance and acceptable-use standards for their adoption. Serve in contract administrator/manager/supervisor and emergency duty roles as assigned; perform other related duties as assigned. These essential job functions are not a complete statement of duties. Employees may be required to perform other related duties as assigned. Requirements Education and Experience Bachelor's degree in Information Technology, Computer Science, Information Systems, Cybersecurity, Business Technology, or a related field. Seven (7) to ten (10) years of progressively responsible IT experience, including leadership over cybersecurity, SaaS/cloud platforms, telecommunications, end user support, vendor management, compliance, and strategic technology planning. Demonstrated experience reading and interpreting SOC 1 (Type II) and SOC 2 (Type II) audit reports, evaluating bridge letters, defining CUECs, and conducting material weakness assessments in a compliance context. Experience leading information security initiatives, technology projects, and business continuity/transformation efforts preferred. Experience in nonprofit, education, healthcare, human services, government, or similarly regulated environments preferred. (A comparable amount of training, education, or experience can be substituted for minimum qualifications.) Skills, Knowledge and Abilities Ability to provide strategic, forward-looking IT leadership, identifying risks and recommending proactive improvements. Demonstrated ability to operate effectively at both the strategic and operational level - equally comfortable presenting cybersecurity risk to the Board and troubleshooting a network outage or SaaS misconfiguration alongside IT staff. Ability to deine problems, interpret technical and diagrammatic instructions, and draw valid conclusions from data click apply for full job details
At Sword, we're building AI to heal billions and unlock humanity's full potential. In doing so, we're pioneering AI Care, a fundamentally new approach to healthcare built for medical reasoning, safety, and real-time treatment, not generic technology applied after the fact. As both a clinical-centric frontier AI lab and an applied AI platform, Sword is reimagining how care is delivered at scale, removing traditional barriers like appointments, waiting rooms, and stigma so more people can access the care they need-and ultimately get back to lives lived in full. Since 2020, Sword has expanded across physical therapy, women's health, cardiometabolic, and mental health, and is now moving beyond the session to a fully AI-native, 24/7 care program that brings physical activity, therapeutic exercise, psychotherapy, nutrition, and behavior change into one connected experience. More than 700,000 members across three continents have completed over 10 million AI sessions, helping 1,000+ enterprise clients avoid more than $1 billion in unnecessary healthcare costs. Backed by 42 clinical studies, 44+ patents, and more than $500 million raised from leading investors including Khosla Ventures, General Catalyst, and Founders Fund, Sword is defining a new standard for healthcare. AI Proficiency at Sword AI fluency is a core expectation at Sword. Every candidate is assessed against our three-level framework - be ready to share real examples of how AI is already part of how you work. Explorer (Level 1) - Uses AI daily to boost personal productivity Builder (Level 2) - Creates workflows and tools that elevate the whole team Integrator (Level 3) - Embeds AI into products and processes at scale Every hire must demonstrate at least Level 1. The expected level will vary depending on the seniority of the role. Role As a Senior DevOps Engineer at Sword Health, you'll own and evolve the infrastructure that powers the world's leading AI Care platform. Working across a multi-cloud environment, you'll build and maintain the systems behind our frontends, backends, microservices, and data pipelines - collaborating closely with multiple engineering teams to keep everything reliable, scalable, and fast. You'll also interface with our AI teams as their models move into production, ensuring the infrastructure is ready to support them. If you thrive in cross-team environments and want your infrastructure work to directly impact healthcare at scale - we'd love to hear from you. To get to know more about our Tech Stack, check here. What you'll be doing Design, implement, and maintain scalable, resilient infrastructure to support Sword Health's high-demand applications and services. Automate and streamline deployment processes, CI/CD pipelines, and routine maintenance tasks to enhance efficiency and reduce downtime. Monitor and optimize system performance, proactively identifying and resolving issues to ensure high availability and reliability. Collaborate closely with development, data, and security teams to ensure seamless integration of infrastructure and code changes. Drive security best practices by implementing and managing access control, network security, and compliance-related policies across the infrastructure. Lead incident response and troubleshooting for infrastructure-related issues, ensuring rapid and effective resolution to maintain service continuity. Mentor and guide junior team members, sharing DevOps best practices and fostering a culture of continuous learning and improvement within the team. Stay up-to-date with industry trends and emerging technologies, bringing innovative solutions to Sword Health's DevOps processes and toolchains. What you need to have Experience with Linux environments. Experience with DevOps and GitOps methodologies. Experience with Kubernetes and Containerized applications (Docker). Experience with Infrastructure as Code (Terraform). Experience with Monitoring Tools (Google Cloud Monitoring/StackDriver, Grafana, Prometheus/AlertManager, NewRelic). Experience with Jenkins. Experience with CI/CD. Team player, Solution-oriented, Proactive attitude with "Get Things Done" mindset. Enthusiast and interested in technologies and innovation. Fluent in English (written and oral). Extra: Experience/Knowledge with Kafka, Prometheus/AlertManager, Grafana, Elasticsearch/ Logstash/ Kibana, Vault, Redis, MySQL, DNS. Usage of AI to debug and develop Infrastructure tooling. Development of AI Agents to automate processes and Infrastructure monitoring and provisioning. Extra: Experience with PHP, Javascript, GoLang. Extra: Experience provisioning servers and services using AWS, Azure, or GCP. Extra: Experience/Knowledge with Istio. Extra: Good know-how about Cloud Networking including VPC Management, Routing, NAT, and overall troubleshooting using TCPdump analysis. The range below reflects base, variable, and equity for this role. This is a starting point, not a ceiling - once someone joins and proves they're outlier talent, we adjust quickly to make sure their compensation matches their impact. Job titles may span more than one career level, so actual pay depends on skills, qualifications, experience, location, and market demand, among other factors. The range reflects base salary plus any variable, bonus, or sales incentives, and our estimate of the value of private company stock options where applicable. It's subject to change, and future stock value isn't guaranteed. Sword also offers additional benefits beyond total compensation. Total Compensation Range $140,000-$220,000 USD Country-specific benefits may vary and will be detailed separately. Sword Health complies with applicable Federal and State civil rights laws and does not discriminate on the basis of Age, Ancestry, Color, Citizenship, Gender, Gender expression, Gender identity, Gender information, Marital status, Medical condition, National origin, Physical or mental disability, Pregnancy, Race, Religion, Caste, Sexual orientation, and Veteran status.
09/23/2026
Full time
At Sword, we're building AI to heal billions and unlock humanity's full potential. In doing so, we're pioneering AI Care, a fundamentally new approach to healthcare built for medical reasoning, safety, and real-time treatment, not generic technology applied after the fact. As both a clinical-centric frontier AI lab and an applied AI platform, Sword is reimagining how care is delivered at scale, removing traditional barriers like appointments, waiting rooms, and stigma so more people can access the care they need-and ultimately get back to lives lived in full. Since 2020, Sword has expanded across physical therapy, women's health, cardiometabolic, and mental health, and is now moving beyond the session to a fully AI-native, 24/7 care program that brings physical activity, therapeutic exercise, psychotherapy, nutrition, and behavior change into one connected experience. More than 700,000 members across three continents have completed over 10 million AI sessions, helping 1,000+ enterprise clients avoid more than $1 billion in unnecessary healthcare costs. Backed by 42 clinical studies, 44+ patents, and more than $500 million raised from leading investors including Khosla Ventures, General Catalyst, and Founders Fund, Sword is defining a new standard for healthcare. AI Proficiency at Sword AI fluency is a core expectation at Sword. Every candidate is assessed against our three-level framework - be ready to share real examples of how AI is already part of how you work. Explorer (Level 1) - Uses AI daily to boost personal productivity Builder (Level 2) - Creates workflows and tools that elevate the whole team Integrator (Level 3) - Embeds AI into products and processes at scale Every hire must demonstrate at least Level 1. The expected level will vary depending on the seniority of the role. Role As a Senior DevOps Engineer at Sword Health, you'll own and evolve the infrastructure that powers the world's leading AI Care platform. Working across a multi-cloud environment, you'll build and maintain the systems behind our frontends, backends, microservices, and data pipelines - collaborating closely with multiple engineering teams to keep everything reliable, scalable, and fast. You'll also interface with our AI teams as their models move into production, ensuring the infrastructure is ready to support them. If you thrive in cross-team environments and want your infrastructure work to directly impact healthcare at scale - we'd love to hear from you. To get to know more about our Tech Stack, check here. What you'll be doing Design, implement, and maintain scalable, resilient infrastructure to support Sword Health's high-demand applications and services. Automate and streamline deployment processes, CI/CD pipelines, and routine maintenance tasks to enhance efficiency and reduce downtime. Monitor and optimize system performance, proactively identifying and resolving issues to ensure high availability and reliability. Collaborate closely with development, data, and security teams to ensure seamless integration of infrastructure and code changes. Drive security best practices by implementing and managing access control, network security, and compliance-related policies across the infrastructure. Lead incident response and troubleshooting for infrastructure-related issues, ensuring rapid and effective resolution to maintain service continuity. Mentor and guide junior team members, sharing DevOps best practices and fostering a culture of continuous learning and improvement within the team. Stay up-to-date with industry trends and emerging technologies, bringing innovative solutions to Sword Health's DevOps processes and toolchains. What you need to have Experience with Linux environments. Experience with DevOps and GitOps methodologies. Experience with Kubernetes and Containerized applications (Docker). Experience with Infrastructure as Code (Terraform). Experience with Monitoring Tools (Google Cloud Monitoring/StackDriver, Grafana, Prometheus/AlertManager, NewRelic). Experience with Jenkins. Experience with CI/CD. Team player, Solution-oriented, Proactive attitude with "Get Things Done" mindset. Enthusiast and interested in technologies and innovation. Fluent in English (written and oral). Extra: Experience/Knowledge with Kafka, Prometheus/AlertManager, Grafana, Elasticsearch/ Logstash/ Kibana, Vault, Redis, MySQL, DNS. Usage of AI to debug and develop Infrastructure tooling. Development of AI Agents to automate processes and Infrastructure monitoring and provisioning. Extra: Experience with PHP, Javascript, GoLang. Extra: Experience provisioning servers and services using AWS, Azure, or GCP. Extra: Experience/Knowledge with Istio. Extra: Good know-how about Cloud Networking including VPC Management, Routing, NAT, and overall troubleshooting using TCPdump analysis. The range below reflects base, variable, and equity for this role. This is a starting point, not a ceiling - once someone joins and proves they're outlier talent, we adjust quickly to make sure their compensation matches their impact. Job titles may span more than one career level, so actual pay depends on skills, qualifications, experience, location, and market demand, among other factors. The range reflects base salary plus any variable, bonus, or sales incentives, and our estimate of the value of private company stock options where applicable. It's subject to change, and future stock value isn't guaranteed. Sword also offers additional benefits beyond total compensation. Total Compensation Range $140,000-$220,000 USD Country-specific benefits may vary and will be detailed separately. Sword Health complies with applicable Federal and State civil rights laws and does not discriminate on the basis of Age, Ancestry, Color, Citizenship, Gender, Gender expression, Gender identity, Gender information, Marital status, Medical condition, National origin, Physical or mental disability, Pregnancy, Race, Religion, Caste, Sexual orientation, and Veteran status.
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
09/23/2026
Full time
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
09/23/2026
Full time
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details