it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

15 jobs found

Email me jobs like this
Refine Search
Current Search
software devops engineer networking
Boeing
Site Reliability Engineer (Associate, Experienced, or Senior)
Boeing Saint Louis, Missouri
Job Description At Boeing, we innovate and collaborate to make the world a better place. We're committed to fostering an environment for every teammate that's welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us. The Boeing Company is looking for a Site Reliability Engineer (Associate, Experienced or Senior) to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO. We are seeking a highly talented, motivated, and creative individual to operate, improve, and sustain mission-critical developer platforms used by Air Dominance engineering teams. This role will provide hands-on technical ownership for GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related software delivery tools such as Artifactory and SonarQube. The selected candidate will drive reliability improvements, automate operational workflows, troubleshoot complex incidents, lead planned maintenance activities, and help establish mature Site Reliability Engineering practices for the team. Our teams are currently hiring for a broad range of experience levels including Associate, Experienced and/or Senior Level Software Engineers. Position Responsibilities: Operate and maintain GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related developer tooling infrastructure Support GitLab runner registration, runner health checks, runner queue troubleshooting, and basic capacity, KPI, and error budget reporting Support software development tool administration, maintenance, version upgrades, patch management, and integration between tools such as Jira, GitLab, Artifactory, Confluence, and SonarQube Serve as a technical owner for platform reliability, availability, performance, capacity, backup, recovery, and operational readiness Develop and maintain Infrastructure as Code (IaC), Ansible, and other automation for provisioning, configuration, platform scaling, health checks, reporting, backup validation, and routine operational tasks Plan and execute approved changes, including application upgrades, security patches, database maintenance, runner lifecycle activities, and infrastructure updates Define, collect, analyze, and refine software delivery and platform reliability metrics to support data-driven decision making Support incident response, root cause analysis, corrective action tracking, and post-incident reviews Partner with developers, project administrators, cybersecurity personnel, infrastructure teams, database administrators, and program stakeholders Improve runbooks, standard operating procedures, architecture documentation, and disaster recovery procedures Evaluate platform risks, capacity trends, recurring incidents, and operational toil, then recommend and implement improvements Participate in after-hours support for urgent or mission-impacting issues as required Monitor application, runner, database, storage, and host health using approved monitoring and alerting tools Triage and resolve routine service requests, access issues, pipeline infrastructure issues, and platform support tickets Assist with incident response during primary support hours and participate in after-hours support when required by mission need Help maintain operational runbooks, troubleshooting guides, architecture notes, and standard operating procedures Assist with backup monitoring, restore validation, patching, upgrades, and planned maintenance activities Create and maintain Infrastructure as Code (IaC), Ansible, and other scripts and automation to simplify infrastructure administration and software deployment under the guidance of more senior engineers Assist in setting up and maintaining development and production-like environments for developer tools and supporting application devices Contribute to metrics and dashboards that monitor system performance, software delivery health, and operational efficiency Follow approved change management, security, access control, and configuration management processes Collaborate with developers, project administrators, cybersecurity personnel, infrastructure teams, and program stakeholders Learn and apply Site Reliability Engineering practices, including incident management, service objectives, root cause analysis, automation, and continuous improvement Contribute to process improvements that help operationally field higher-quality end-to-end system software more frequently This position is expected to be 100% onsite. The selected candidate will be required to work onsite at one of the listed location options. Basic Qualifications (Required Skills/ Experience): Bachelor's Degree This position requires the ability to obtain a US Security Clearance for which the US Government requires US Citizenship as a condition of employment (An interim and/or final U.S. Secret Clearance Post-Start may be required) This position requires the ability to obtain access to Special Access Programs (SAP), for which the US Government requires US Citizenship as a condition of employment 2+ years of experience with software development and/or troubleshooting software 2+ years of experience with Git-based source control workflows including branching strategies, code reviews, and pull request processes 2+ years of experience developing software products in a cloud computing environment e.g. Azure/AWS/Google Cloud Preferred Qualifications (Desired Skills/Experience): Level 3: 5 or more years' related work experience or an equivalent combination of education and experience Level 4: 9 or more years' related work experience or an equivalent combination of education and experience Active clearance Linux system administration, software development, DevOps, DevSecOps, IT operations, or related technical work Experience or coursework with Agile software development Basic understanding of networking, operating systems, databases, software build processes, and secure system administration Ability to follow documented procedures and communicate technical status clearly Experience with Jira, Confluence, or other Atlassian administration and support activities Experience with PostgreSQL administration, SQL troubleshooting, backups, or restore procedures Experience writing scripts in Bash, Python, PowerShell, Java, C#, C++, or a similar language Experience with Artifactory, SonarQube, Jenkins, or similar software delivery tools Familiarity with monitoring, alerting, logging, incident response, and operational metrics Ability to obtain Security+ certification Experience designing or improving monitoring, alerting, dashboards, operational metrics, and SLOs Experience supporting Air Dominance, classified, air-gapped, or highly regulated engineering environments Experience with disaster recovery planning, continuity of operations, backup validation, and restore testing Experience leading technical investigations, post-incident reviews, and corrective action plans Demonstrated ability to balance customer urgency, system reliability, security requirements, and change control Strong customer service mindset and ability to support engineering teams in a mission-focused environment Strong written and verbal communication skills Travel: 5% Drug Free Workplace: Boeing is a Drug Free Workplace (DFW) where post offer applicants and employees are subject to testing for marijuana, cocaine, opioids, amphetamines, PCP, and alcohol when criteria is met as outlined in our policies. Conflict of Interest: Successful candidates for this job must satisfy the Company's Conflict of Interest (COI) assessment process. CodeVue Coding Challenge: To be considered for this position you will be required to complete a technical assessment as part of the selection process. Failure to complete the assessment will remove you from consideration. Pay & Benefits: At Boeing, we strive to deliver a Total Rewards package that will attract, engage and retain the top talent. Elements of the Total Rewards package include competitive base pay and variable compensation opportunities. The Boeing Company also provides eligible employees with an opportunity to enroll in a variety of benefit programs, generally including health insurance, flexible spending accounts, health savings accounts, retirement savings plans, life and disability insurance programs, and a number of programs that provide for both paid and unpaid time away from work. The specific programs and options available to any given employee may vary depending on eligibility factors such as geographic location, date of hire, and the applicability of collective bargaining agreements. Pay is based upon candidate experience and qualifications, as well as market and business considerations. Summary Pay Range for Associate level (Level 2): $99,450 - $134,550 Summary Pay Range for Experienced level (Level 3): $126,650 - $171,350 Summary Pay Range for Senior level (Level 4): $160,650 - $217,350 Applications for this position will be accepted until Oct. 01, 2026 Export Control Requirements: This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a "U.S . click apply for full job details
09/21/2026
Full time
Job Description At Boeing, we innovate and collaborate to make the world a better place. We're committed to fostering an environment for every teammate that's welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us. The Boeing Company is looking for a Site Reliability Engineer (Associate, Experienced or Senior) to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO. We are seeking a highly talented, motivated, and creative individual to operate, improve, and sustain mission-critical developer platforms used by Air Dominance engineering teams. This role will provide hands-on technical ownership for GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related software delivery tools such as Artifactory and SonarQube. The selected candidate will drive reliability improvements, automate operational workflows, troubleshoot complex incidents, lead planned maintenance activities, and help establish mature Site Reliability Engineering practices for the team. Our teams are currently hiring for a broad range of experience levels including Associate, Experienced and/or Senior Level Software Engineers. Position Responsibilities: Operate and maintain GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related developer tooling infrastructure Support GitLab runner registration, runner health checks, runner queue troubleshooting, and basic capacity, KPI, and error budget reporting Support software development tool administration, maintenance, version upgrades, patch management, and integration between tools such as Jira, GitLab, Artifactory, Confluence, and SonarQube Serve as a technical owner for platform reliability, availability, performance, capacity, backup, recovery, and operational readiness Develop and maintain Infrastructure as Code (IaC), Ansible, and other automation for provisioning, configuration, platform scaling, health checks, reporting, backup validation, and routine operational tasks Plan and execute approved changes, including application upgrades, security patches, database maintenance, runner lifecycle activities, and infrastructure updates Define, collect, analyze, and refine software delivery and platform reliability metrics to support data-driven decision making Support incident response, root cause analysis, corrective action tracking, and post-incident reviews Partner with developers, project administrators, cybersecurity personnel, infrastructure teams, database administrators, and program stakeholders Improve runbooks, standard operating procedures, architecture documentation, and disaster recovery procedures Evaluate platform risks, capacity trends, recurring incidents, and operational toil, then recommend and implement improvements Participate in after-hours support for urgent or mission-impacting issues as required Monitor application, runner, database, storage, and host health using approved monitoring and alerting tools Triage and resolve routine service requests, access issues, pipeline infrastructure issues, and platform support tickets Assist with incident response during primary support hours and participate in after-hours support when required by mission need Help maintain operational runbooks, troubleshooting guides, architecture notes, and standard operating procedures Assist with backup monitoring, restore validation, patching, upgrades, and planned maintenance activities Create and maintain Infrastructure as Code (IaC), Ansible, and other scripts and automation to simplify infrastructure administration and software deployment under the guidance of more senior engineers Assist in setting up and maintaining development and production-like environments for developer tools and supporting application devices Contribute to metrics and dashboards that monitor system performance, software delivery health, and operational efficiency Follow approved change management, security, access control, and configuration management processes Collaborate with developers, project administrators, cybersecurity personnel, infrastructure teams, and program stakeholders Learn and apply Site Reliability Engineering practices, including incident management, service objectives, root cause analysis, automation, and continuous improvement Contribute to process improvements that help operationally field higher-quality end-to-end system software more frequently This position is expected to be 100% onsite. The selected candidate will be required to work onsite at one of the listed location options. Basic Qualifications (Required Skills/ Experience): Bachelor's Degree This position requires the ability to obtain a US Security Clearance for which the US Government requires US Citizenship as a condition of employment (An interim and/or final U.S. Secret Clearance Post-Start may be required) This position requires the ability to obtain access to Special Access Programs (SAP), for which the US Government requires US Citizenship as a condition of employment 2+ years of experience with software development and/or troubleshooting software 2+ years of experience with Git-based source control workflows including branching strategies, code reviews, and pull request processes 2+ years of experience developing software products in a cloud computing environment e.g. Azure/AWS/Google Cloud Preferred Qualifications (Desired Skills/Experience): Level 3: 5 or more years' related work experience or an equivalent combination of education and experience Level 4: 9 or more years' related work experience or an equivalent combination of education and experience Active clearance Linux system administration, software development, DevOps, DevSecOps, IT operations, or related technical work Experience or coursework with Agile software development Basic understanding of networking, operating systems, databases, software build processes, and secure system administration Ability to follow documented procedures and communicate technical status clearly Experience with Jira, Confluence, or other Atlassian administration and support activities Experience with PostgreSQL administration, SQL troubleshooting, backups, or restore procedures Experience writing scripts in Bash, Python, PowerShell, Java, C#, C++, or a similar language Experience with Artifactory, SonarQube, Jenkins, or similar software delivery tools Familiarity with monitoring, alerting, logging, incident response, and operational metrics Ability to obtain Security+ certification Experience designing or improving monitoring, alerting, dashboards, operational metrics, and SLOs Experience supporting Air Dominance, classified, air-gapped, or highly regulated engineering environments Experience with disaster recovery planning, continuity of operations, backup validation, and restore testing Experience leading technical investigations, post-incident reviews, and corrective action plans Demonstrated ability to balance customer urgency, system reliability, security requirements, and change control Strong customer service mindset and ability to support engineering teams in a mission-focused environment Strong written and verbal communication skills Travel: 5% Drug Free Workplace: Boeing is a Drug Free Workplace (DFW) where post offer applicants and employees are subject to testing for marijuana, cocaine, opioids, amphetamines, PCP, and alcohol when criteria is met as outlined in our policies. Conflict of Interest: Successful candidates for this job must satisfy the Company's Conflict of Interest (COI) assessment process. CodeVue Coding Challenge: To be considered for this position you will be required to complete a technical assessment as part of the selection process. Failure to complete the assessment will remove you from consideration. Pay & Benefits: At Boeing, we strive to deliver a Total Rewards package that will attract, engage and retain the top talent. Elements of the Total Rewards package include competitive base pay and variable compensation opportunities. The Boeing Company also provides eligible employees with an opportunity to enroll in a variety of benefit programs, generally including health insurance, flexible spending accounts, health savings accounts, retirement savings plans, life and disability insurance programs, and a number of programs that provide for both paid and unpaid time away from work. The specific programs and options available to any given employee may vary depending on eligibility factors such as geographic location, date of hire, and the applicability of collective bargaining agreements. Pay is based upon candidate experience and qualifications, as well as market and business considerations. Summary Pay Range for Associate level (Level 2): $99,450 - $134,550 Summary Pay Range for Experienced level (Level 3): $126,650 - $171,350 Summary Pay Range for Senior level (Level 4): $160,650 - $217,350 Applications for this position will be accepted until Oct. 01, 2026 Export Control Requirements: This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a "U.S . click apply for full job details
Solutions Architect - EA (Azure)
Code Plus Inc Huntsville, Alabama
Job Description Job Description Why CODEplus CODEplus is a software engineering and modernization company with over 31 years supporting federal agencies including DoD, NASA, DOE, NRC, and USPS. We combine the agility and innovation of a small business with mature engineering, Agile, and DevSecOps delivery practices to support mission-critical modernization efforts. Position Title: Solutions Architect - EA (Azure) Various position levels for each position: Junior-level - 2-3 years of experience Mid-level - 4-8 years of experience Senior-level - 10+ years of experience Location: Huntsville, AL (Hybrid/On-Site preferred, possibly remote) Clearance: Active Secret Clearance Required Position Overview CODEplus is seeking a Solutions Architect - EA (Azure) to design, lead, and execute large-scale cloud modernization efforts for Department of Defense and federal customers. This role is ideal for a hands-on cloud architect with deep engineering expertise , who has led enterprise migrations, designed secure cloud platforms, and advised senior stakeholders on cloud strategy and transformation . You will serve as the technical authority for Azure architecture , driving solution design across complex, mission-critical environments , while actively contributing to implementation alongside Agile and DevSecOps teams. This position blends architecture leadership, hands-on engineering, and customer advisory , consistent with CODEplus's engineering-first culture. Key Responsibilities Azure Architecture & Platform Design Lead architecture and design of enterprise-scale Azure solutions , including Azure Government and hybrid environments Design and implement Azure landing zones , multi-region architectures, and high-availability solutions Define reusable cloud patterns, reference architectures, and platform standards Architect solutions aligned with Well-Architected Framework principles (Azure CAF) , performance, cost optimization, and operational excellence Ensure compliance with DoD security requirements (IL4/IL5, RMF, Zero Trust, NIST 800-53) Cloud Modernization & Migration Lead large-scale cloud migrations and transformations , including legacy and cloud-native workloads Architect modernization approaches including: Rehosting, replatforming, and refactoring strategies Containerization and microservices (AKS) Serverless and event-driven architectures Support migration planning for 100+ application portfolios , including dependency mapping and sequencing Hands-on Engineering & DevSecOps Actively contribute to implementation, including: Infrastructure as Code ( Terraform, Bicep, ARM ) CI/CD pipelines and DevSecOps automation Security automation and compliance enforcement Develop or guide development of automation scripts (Python, PowerShell) for infrastructure and operations Lead design reviews, code reviews, and ensure architectural integrity across deployments Agile Leadership & Technical Execution Provide technical leadership to Agile/SAFe teams , supporting PI planning, backlog refinement, and sprint execution Mentor engineers in cloud engineering, DevOps, and architecture best practices Resolve complex system-level challenges across performance, scalability, and security Customer Advisory & Stakeholder Engagement Serve as a trusted advisor to government customers and senior leadership Translate mission needs into technical strategy and cloud architecture decisions Deliver: Executive briefings Architecture diagrams Trade studies and decision papers Participate in TIMs, design reviews, and program governance Integration & Systems Engineering Lead integration across: Cloud services On-prem / hybrid environments External mission systems Collaborate across cybersecurity, networking, software, and systems engineering teams Cost, Risk & Program Support Support cloud cost modeling, optimization strategies , and trade-off analysis Identify technical risks and drive mitigation strategies Contribute to proposals, including: Architecture approaches Basis of Estimate (BOE) Technical volumes Required Qualifications Bachelor's degree in STEM Deep, hands-on Azure architecture experience , including: Networking, identity (Entra ID), compute, storage, and security Azure Government or regulated cloud environments Demonstrated experience: Leading enterprise-scale cloud migrations or transformations Designing solutions across multi-region, high-availability environments Strong hands-on experience with: Infrastructure as Code (Terraform, Bicep, ARM) CI/CD and DevSecOps pipelines Automation scripting ( Python, PowerShell ) Proven ability to: Lead technical teams Work directly with government customers Balance technical excellence, cost, and delivery timelines Strong communication skills with experience briefing senior stakeholders Preferred Qualifications Azure certifications: Azure Solutions Architect Expert Azure Security Engineer Azure DevOps Engineer Experience with: Azure Landing Zones / CAF implementation Containers and Kubernetes ( AKS ) Serverless (Functions, Event Grid, Service Bus) RMF / NIST 800-53 / FedRAMP Platform engineering and internal developer platforms Prior experience supporting DoD, IC, or federal programs Background in systems engineering, DevOps, or software engineering Experience with large-scale migrations (100+ workloads) Experience contributing to capture, proposals, or re-competes
09/20/2026
Full time
Job Description Job Description Why CODEplus CODEplus is a software engineering and modernization company with over 31 years supporting federal agencies including DoD, NASA, DOE, NRC, and USPS. We combine the agility and innovation of a small business with mature engineering, Agile, and DevSecOps delivery practices to support mission-critical modernization efforts. Position Title: Solutions Architect - EA (Azure) Various position levels for each position: Junior-level - 2-3 years of experience Mid-level - 4-8 years of experience Senior-level - 10+ years of experience Location: Huntsville, AL (Hybrid/On-Site preferred, possibly remote) Clearance: Active Secret Clearance Required Position Overview CODEplus is seeking a Solutions Architect - EA (Azure) to design, lead, and execute large-scale cloud modernization efforts for Department of Defense and federal customers. This role is ideal for a hands-on cloud architect with deep engineering expertise , who has led enterprise migrations, designed secure cloud platforms, and advised senior stakeholders on cloud strategy and transformation . You will serve as the technical authority for Azure architecture , driving solution design across complex, mission-critical environments , while actively contributing to implementation alongside Agile and DevSecOps teams. This position blends architecture leadership, hands-on engineering, and customer advisory , consistent with CODEplus's engineering-first culture. Key Responsibilities Azure Architecture & Platform Design Lead architecture and design of enterprise-scale Azure solutions , including Azure Government and hybrid environments Design and implement Azure landing zones , multi-region architectures, and high-availability solutions Define reusable cloud patterns, reference architectures, and platform standards Architect solutions aligned with Well-Architected Framework principles (Azure CAF) , performance, cost optimization, and operational excellence Ensure compliance with DoD security requirements (IL4/IL5, RMF, Zero Trust, NIST 800-53) Cloud Modernization & Migration Lead large-scale cloud migrations and transformations , including legacy and cloud-native workloads Architect modernization approaches including: Rehosting, replatforming, and refactoring strategies Containerization and microservices (AKS) Serverless and event-driven architectures Support migration planning for 100+ application portfolios , including dependency mapping and sequencing Hands-on Engineering & DevSecOps Actively contribute to implementation, including: Infrastructure as Code ( Terraform, Bicep, ARM ) CI/CD pipelines and DevSecOps automation Security automation and compliance enforcement Develop or guide development of automation scripts (Python, PowerShell) for infrastructure and operations Lead design reviews, code reviews, and ensure architectural integrity across deployments Agile Leadership & Technical Execution Provide technical leadership to Agile/SAFe teams , supporting PI planning, backlog refinement, and sprint execution Mentor engineers in cloud engineering, DevOps, and architecture best practices Resolve complex system-level challenges across performance, scalability, and security Customer Advisory & Stakeholder Engagement Serve as a trusted advisor to government customers and senior leadership Translate mission needs into technical strategy and cloud architecture decisions Deliver: Executive briefings Architecture diagrams Trade studies and decision papers Participate in TIMs, design reviews, and program governance Integration & Systems Engineering Lead integration across: Cloud services On-prem / hybrid environments External mission systems Collaborate across cybersecurity, networking, software, and systems engineering teams Cost, Risk & Program Support Support cloud cost modeling, optimization strategies , and trade-off analysis Identify technical risks and drive mitigation strategies Contribute to proposals, including: Architecture approaches Basis of Estimate (BOE) Technical volumes Required Qualifications Bachelor's degree in STEM Deep, hands-on Azure architecture experience , including: Networking, identity (Entra ID), compute, storage, and security Azure Government or regulated cloud environments Demonstrated experience: Leading enterprise-scale cloud migrations or transformations Designing solutions across multi-region, high-availability environments Strong hands-on experience with: Infrastructure as Code (Terraform, Bicep, ARM) CI/CD and DevSecOps pipelines Automation scripting ( Python, PowerShell ) Proven ability to: Lead technical teams Work directly with government customers Balance technical excellence, cost, and delivery timelines Strong communication skills with experience briefing senior stakeholders Preferred Qualifications Azure certifications: Azure Solutions Architect Expert Azure Security Engineer Azure DevOps Engineer Experience with: Azure Landing Zones / CAF implementation Containers and Kubernetes ( AKS ) Serverless (Functions, Event Grid, Service Bus) RMF / NIST 800-53 / FedRAMP Platform engineering and internal developer platforms Prior experience supporting DoD, IC, or federal programs Background in systems engineering, DevOps, or software engineering Experience with large-scale migrations (100+ workloads) Experience contributing to capture, proposals, or re-competes
Technical Lead Market Data
BloomGuarden New York, New York
Job Description Job Description Millennium NYC 5 days a week in office 200-225K Base plus bonus (total comp-400-500K) Technical Lead Market Data The SPEED Market Data team seeks a hands-on Technical Lead who will own and drive a critical workstream focused on architecting, implementing, monitoring, and supporting low-latency C++ systems that are robust, resilient, accurate, stable, and blindingly fast. By leading the design and evolution of this high-performance infrastructure, you will help position MLP as a leader in quantitative trading. You will shape the future of this industry while working alongside other exceptional engineers and strategists to solve some of the most significant engineering problems in the world. We are looking for a strong technical leader with financial markets technology experience and real-time market data expertise to design, build, and support our global real-time (both low-latency and non-latency-sensitive) market data platform. This role emphasizes technical leadership, architectural ownership, and cross-team coordination rather than people management. The successful candidate will be comfortable owning a workstream end-to-end-covering design, implementation, monitoring, support, and stakeholder management-while ensuring stability of the existing environment and driving platform improvements. Principal Responsibilities Act as the technical owner for a major market data workstream, setting technical direction, defining architecture, and driving execution across the full lifecycle. Collaborate with hardware and software teams across divisions to design and build real-time market data processing and distribution systems. Lead and drive new technical initiatives for the team, including evaluating technologies, defining standards, and establishing best practices. Design and develop systems, interfaces, and tools for historical market data and trading simulations that increase research productivity. Architect and implement components of an enterprise market data platform, including components for caching, aggregation, conflation and value-added data enrichment. Optimize platform performance using network and systems programming, and advanced low-latency techniques (CPU, NIC, kernel, and application-level tuning). Lead the design and maintenance of automated test and benchmark frameworks, and tools for risk management, performance tracking, and system validation. Provide technical leadership for the support and operation of both enterprise real-time market data environments, including coordinating internal, vendor, and exchange-driven changes. Design and engineer components to automate support and management of the market data platform, including monitoring, real-time and historical metrics collection/visualization, and self-service administrative/user tools. Serve as a primary technical liaison for users of the market data environment (Portfolio Managers, trading desks, and core technology teams), translating requirements into robust technical solutions. Lead the enhancement of processes and workflows for operating the market data platform (release/deployment, incident management and remediation, exchange notification handling, defining and enforcing SLAs). Mentor and influence other engineers through code reviews, design reviews, and hands-on guidance, fostering a culture of technical excellence and accountability. Qualifications / Skills Required Degree in Computer Science or a related field with a strong background in data structures, algorithms, and object-oriented programming in modern C++. Deep understanding of Linux system internals and networking, especially in low-latency and high-throughput environments. Strong knowledge of CPU architecture and the ability to leverage CPU capabilities for performance optimization. Demonstrated experience acting as a technical lead or senior engineer owning complex systems or workstreams end-to-end (design, delivery, and operations). Able to prioritize and make trade-offs in a fast-moving, high-pressure, constantly changing environment; strong sense of urgency, ownership, and follow-through. Strong belief in and practice of extreme ownership, with a track record of taking accountability for systems in production. Effective communication and stakeholder management skills: able to work closely with business and technology users, understand their needs, and drive appropriate technical solutions. Experience building solutions on cloud environments such as GCP and AWS. Knowledge of additional programming languages such as Java, Python, or scripting (Perl, shell). Technical background in application development on complex market data systems (e.g., Bloomberg, Thomson Reuters, etc.). Experience supporting market data environments within a global organization, including internally developed DMA feed handlers and distribution infrastructure. Strong understanding of market data concepts and functionality, including data models (fields/messages), protocols (e.g., snapshot + delta), order book representations (L1/L2/L3), recovery, and reliability. Hands-on Site Reliability Engineering or DevOps experience, including system administration, automation, measurement, and release/deployment management. Experience with monitoring, metrics, and command/control tooling for distributed market data platforms, with the ability to evaluate existing solutions and drive enhancements across development and operations. Ability to operate with a high level of thoroughness and attention to detail, demonstrating strong ownership of deliverables and production systems.
09/20/2026
Full time
Job Description Job Description Millennium NYC 5 days a week in office 200-225K Base plus bonus (total comp-400-500K) Technical Lead Market Data The SPEED Market Data team seeks a hands-on Technical Lead who will own and drive a critical workstream focused on architecting, implementing, monitoring, and supporting low-latency C++ systems that are robust, resilient, accurate, stable, and blindingly fast. By leading the design and evolution of this high-performance infrastructure, you will help position MLP as a leader in quantitative trading. You will shape the future of this industry while working alongside other exceptional engineers and strategists to solve some of the most significant engineering problems in the world. We are looking for a strong technical leader with financial markets technology experience and real-time market data expertise to design, build, and support our global real-time (both low-latency and non-latency-sensitive) market data platform. This role emphasizes technical leadership, architectural ownership, and cross-team coordination rather than people management. The successful candidate will be comfortable owning a workstream end-to-end-covering design, implementation, monitoring, support, and stakeholder management-while ensuring stability of the existing environment and driving platform improvements. Principal Responsibilities Act as the technical owner for a major market data workstream, setting technical direction, defining architecture, and driving execution across the full lifecycle. Collaborate with hardware and software teams across divisions to design and build real-time market data processing and distribution systems. Lead and drive new technical initiatives for the team, including evaluating technologies, defining standards, and establishing best practices. Design and develop systems, interfaces, and tools for historical market data and trading simulations that increase research productivity. Architect and implement components of an enterprise market data platform, including components for caching, aggregation, conflation and value-added data enrichment. Optimize platform performance using network and systems programming, and advanced low-latency techniques (CPU, NIC, kernel, and application-level tuning). Lead the design and maintenance of automated test and benchmark frameworks, and tools for risk management, performance tracking, and system validation. Provide technical leadership for the support and operation of both enterprise real-time market data environments, including coordinating internal, vendor, and exchange-driven changes. Design and engineer components to automate support and management of the market data platform, including monitoring, real-time and historical metrics collection/visualization, and self-service administrative/user tools. Serve as a primary technical liaison for users of the market data environment (Portfolio Managers, trading desks, and core technology teams), translating requirements into robust technical solutions. Lead the enhancement of processes and workflows for operating the market data platform (release/deployment, incident management and remediation, exchange notification handling, defining and enforcing SLAs). Mentor and influence other engineers through code reviews, design reviews, and hands-on guidance, fostering a culture of technical excellence and accountability. Qualifications / Skills Required Degree in Computer Science or a related field with a strong background in data structures, algorithms, and object-oriented programming in modern C++. Deep understanding of Linux system internals and networking, especially in low-latency and high-throughput environments. Strong knowledge of CPU architecture and the ability to leverage CPU capabilities for performance optimization. Demonstrated experience acting as a technical lead or senior engineer owning complex systems or workstreams end-to-end (design, delivery, and operations). Able to prioritize and make trade-offs in a fast-moving, high-pressure, constantly changing environment; strong sense of urgency, ownership, and follow-through. Strong belief in and practice of extreme ownership, with a track record of taking accountability for systems in production. Effective communication and stakeholder management skills: able to work closely with business and technology users, understand their needs, and drive appropriate technical solutions. Experience building solutions on cloud environments such as GCP and AWS. Knowledge of additional programming languages such as Java, Python, or scripting (Perl, shell). Technical background in application development on complex market data systems (e.g., Bloomberg, Thomson Reuters, etc.). Experience supporting market data environments within a global organization, including internally developed DMA feed handlers and distribution infrastructure. Strong understanding of market data concepts and functionality, including data models (fields/messages), protocols (e.g., snapshot + delta), order book representations (L1/L2/L3), recovery, and reliability. Hands-on Site Reliability Engineering or DevOps experience, including system administration, automation, measurement, and release/deployment management. Experience with monitoring, metrics, and command/control tooling for distributed market data platforms, with the ability to evaluate existing solutions and drive enhancements across development and operations. Ability to operate with a high level of thoroughness and attention to detail, demonstrating strong ownership of deliverables and production systems.
AI & Emerging Tech Specialist, Emerging Technology
Amazon Web Services, Inc.
Amazon Web Services (AWS) is the leading cloud provider. AWS runs a globally distributed and resilient environment, operating at massive scale and enabling businesses and government agencies to run their operations and applications on AWS's multi-tenant infrastructure. Through AWS' virtual infrastructure, customers all over the world innovate faster with emerging technologies like AI/machine learning including generative AI, high performance computing, quantum, and more. Our World Wide Public Sector (WWPS) team is looking for an exceptional sales candidate who is excited about accelerating customers' journey to transform their mission areas and business functions with emerging technology. The candidate will join our fast growing AWS National Security Emerging Technology team to support sensitive and critical US Defense and Intelligence cloud technologies and activities. In this AI & Emerging Tech Specialist role, you will focus on developing customer use cases, driving adoption, and shaping the go-to-market strategy across multiple customer agencies in close collaboration with each account team. You will work backwards from customers' mission needs to develop integrated solutions that leverage AWS and partner data, analytics, generative AI, and emerging technology capabilities. As an AI & Emerging Tech Specialist, you will partner with customer executives and builders as they adopt DataOps, MLOps, and AI governance best practices to operationalize AI at scale in secure and classified environments. This position requires that the candidate selected be a US Citizen and must currently possess an active Top Secret security clearance. The position further requires that, after start, the selected candidate obtain and maintain an active TS/SCI security clearance with polygraph and satisfy other security related requirements. Key job responsibilities - Prospect into new AI and emerging technology customers and buying centers by working closely with AWS account teams and partners to identify mission-driven use cases for generative AI, agentic AI, ML at the edge, quantum, and other emerging capabilities. - Serve as a subject matter expert in AI and emerging technologies - including generative AI, foundation models, and associated AWS services - providing specialized sales guidance, technical enablement, and thought leadership to account teams and customers. - Meet regularly with national security executives and mission owners to understand their mission requirements, existing architectures, and technology roadmaps, and introduce AWS AI, data, analytics, and emerging technology services that accelerate mission outcomes. - Negotiate and close large, complex AI and emerging technology opportunities, working across multiple agencies and buying centers to drive adoption at scale. - Work closely with Amazon business leaders - including customer account teams, hardware and software engineering leaders, professional services, and technical program managers - to align emerging technology solutions with customer mission needs. - Represent the voice of the customer; collaborate with field and central teams to bring customer feedback to product teams. Lead curation of custom feature requests, regional availability requirements, and unique mission use cases across classified and unclassified environments. - Identify opportunities for Marketing and Public Policy engagements that position AWS as a trusted partner for AI and emerging technologies in the national security community. - Leverage Defense and Professional Associations to establish AWS as a thought and industry leader in AI and emerging technologies, deepening customer intimacy, nurturing existing partner relationships, developing new partnerships, and uncovering new opportunities. A day in the life This is a highly dynamic role, and as such, your typical day could include 2-3 morning meetings with agency executives, program officers, and builders. In the afternoon, you might be leading an internal strategy or working session with key stakeholders from NatSec account teams, service teams, or Professional Services. About the team The team you would join specializes in AI and Emerging Technologies for National Security, with a core focus on generative AI, frontier AI, ML at the edge, quantum, mission networking, and advanced computing. We pride ourselves on our deep technical expertise and mission knowledge. Members of our team dive deep into their technical specialties, stay ahead of trends in mission adoption of emerging technologies, and develop best practices for accelerating responsible adoption across classified and unclassified environments. As a high-performing sales organization embedded within the NatSec community, the team serves as a trusted partner to end users, agency leadership, and the policy community - helping customers navigate the rapidly evolving AI and emerging technology landscape to deliver mission impact. Diverse Experiences Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Why AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, inspire us to never stop embracing our uniqueness. Mentorship and Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS- Bachelor's degree in a relevant field or equivalent work experience - 3+ years of technology platform sales with an understanding of government IT, data centers, cloud services and cloud adoption experience - Experience working cross-functionally with tech and non-tech teams - Current, active US Government Security Clearance of Top Secret or above PREFERRED QUALIFICATIONS- 5+ years of direct sales or business development in software, cloud or SaaS markets selling to C-level executives experience - Experience with one or more of the following domains: analytics, security, storage, DevOps, application development, or machine learning - Experience selling AI/ML solutions - Experience interpreting data and making business recommendations - Experience developing, negotiating and executing business agreements Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, MD, Jessup - 92 000.00 USD annually USA, VA, Arlington - 92 000.00 USD annually USA, VA, Herndon - 92 000.00 USD annually
09/20/2026
Full time
Amazon Web Services (AWS) is the leading cloud provider. AWS runs a globally distributed and resilient environment, operating at massive scale and enabling businesses and government agencies to run their operations and applications on AWS's multi-tenant infrastructure. Through AWS' virtual infrastructure, customers all over the world innovate faster with emerging technologies like AI/machine learning including generative AI, high performance computing, quantum, and more. Our World Wide Public Sector (WWPS) team is looking for an exceptional sales candidate who is excited about accelerating customers' journey to transform their mission areas and business functions with emerging technology. The candidate will join our fast growing AWS National Security Emerging Technology team to support sensitive and critical US Defense and Intelligence cloud technologies and activities. In this AI & Emerging Tech Specialist role, you will focus on developing customer use cases, driving adoption, and shaping the go-to-market strategy across multiple customer agencies in close collaboration with each account team. You will work backwards from customers' mission needs to develop integrated solutions that leverage AWS and partner data, analytics, generative AI, and emerging technology capabilities. As an AI & Emerging Tech Specialist, you will partner with customer executives and builders as they adopt DataOps, MLOps, and AI governance best practices to operationalize AI at scale in secure and classified environments. This position requires that the candidate selected be a US Citizen and must currently possess an active Top Secret security clearance. The position further requires that, after start, the selected candidate obtain and maintain an active TS/SCI security clearance with polygraph and satisfy other security related requirements. Key job responsibilities - Prospect into new AI and emerging technology customers and buying centers by working closely with AWS account teams and partners to identify mission-driven use cases for generative AI, agentic AI, ML at the edge, quantum, and other emerging capabilities. - Serve as a subject matter expert in AI and emerging technologies - including generative AI, foundation models, and associated AWS services - providing specialized sales guidance, technical enablement, and thought leadership to account teams and customers. - Meet regularly with national security executives and mission owners to understand their mission requirements, existing architectures, and technology roadmaps, and introduce AWS AI, data, analytics, and emerging technology services that accelerate mission outcomes. - Negotiate and close large, complex AI and emerging technology opportunities, working across multiple agencies and buying centers to drive adoption at scale. - Work closely with Amazon business leaders - including customer account teams, hardware and software engineering leaders, professional services, and technical program managers - to align emerging technology solutions with customer mission needs. - Represent the voice of the customer; collaborate with field and central teams to bring customer feedback to product teams. Lead curation of custom feature requests, regional availability requirements, and unique mission use cases across classified and unclassified environments. - Identify opportunities for Marketing and Public Policy engagements that position AWS as a trusted partner for AI and emerging technologies in the national security community. - Leverage Defense and Professional Associations to establish AWS as a thought and industry leader in AI and emerging technologies, deepening customer intimacy, nurturing existing partner relationships, developing new partnerships, and uncovering new opportunities. A day in the life This is a highly dynamic role, and as such, your typical day could include 2-3 morning meetings with agency executives, program officers, and builders. In the afternoon, you might be leading an internal strategy or working session with key stakeholders from NatSec account teams, service teams, or Professional Services. About the team The team you would join specializes in AI and Emerging Technologies for National Security, with a core focus on generative AI, frontier AI, ML at the edge, quantum, mission networking, and advanced computing. We pride ourselves on our deep technical expertise and mission knowledge. Members of our team dive deep into their technical specialties, stay ahead of trends in mission adoption of emerging technologies, and develop best practices for accelerating responsible adoption across classified and unclassified environments. As a high-performing sales organization embedded within the NatSec community, the team serves as a trusted partner to end users, agency leadership, and the policy community - helping customers navigate the rapidly evolving AI and emerging technology landscape to deliver mission impact. Diverse Experiences Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Why AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, inspire us to never stop embracing our uniqueness. Mentorship and Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS- Bachelor's degree in a relevant field or equivalent work experience - 3+ years of technology platform sales with an understanding of government IT, data centers, cloud services and cloud adoption experience - Experience working cross-functionally with tech and non-tech teams - Current, active US Government Security Clearance of Top Secret or above PREFERRED QUALIFICATIONS- 5+ years of direct sales or business development in software, cloud or SaaS markets selling to C-level executives experience - Experience with one or more of the following domains: analytics, security, storage, DevOps, application development, or machine learning - Experience selling AI/ML solutions - Experience interpreting data and making business recommendations - Experience developing, negotiating and executing business agreements Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, MD, Jessup - 92 000.00 USD annually USA, VA, Arlington - 92 000.00 USD annually USA, VA, Herndon - 92 000.00 USD annually
SDE II, Neuron Infra Services
Annapurna Labs (U.S.) Inc. Cupertino, California
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the trn and inf servers that use them. This position is for a Software Engineer that will lead the development of various services that will aid in optimization, analysis and release of machine learning workloads and artifacts. This candidate must have had experience leading distributed systems and machine learning related projects, preferably starting from architecture through several generations of delivery to customers. Deep knowledge of optimization, resource management, scheduling are needed. The ideal candidate will have experience working on services like EC2, EKS, Lambda in AWS or similar services on other cloud providers. Key job responsibilities This engineer will lead the design and implementation of new tools, pipelines and automation, will work with developers, system architects, hardware engineers and users both within and external to Amazon to ensure compatibility of this new toolset with existing and next-generation AI accelerators. Design, implement, and maintain CI/CD pipelines to automate the software release process. Collaborate with development teams to integrate new software releases. Infrastructure Management: Manage and automate infrastructure provisioning. Ensure high availability and scalability of systems through effective infrastructure management. Monitoring and Optimization: Implement monitoring solutions to track system performance. Identify bottlenecks and optimize system performance. Security and Compliance: Implement security best practices in the DevOps pipeline. Conduct regular vulnerability assessments and risk management. A day in the life As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also: Build high-impact solutions to deliver to our large customer base. Participate in design discussions, code review, and communicate with internal and external stakeholders. Work cross-functionally to help drive business decisions with your technical input. Work in a startup-like development environment, where you're always working on the most important stuff. About the team The Neuron Infra Services team fosters a builder's culture where experimentation is encouraged, and impact is measurable. We emphasize collaboration, technical ownership, and continuous learning. Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Knowledge of system performance, memory management, and parallel computing principles - Experience in debugging, profiling, and implementing software engineering best practices in large-scale systems, or experience debugging, profiling, and implementing best software engineering practices in large-scale systems - Experience with AWS or cloud technologies PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - Knowledge of fundamentals of networking, security, databases (relational or NoSQL), operating systems (Unix, Linux, and/or Windows) - Fundamentals of Machine learning and LLMs, their architecture along with work experience on certain LLM models. Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the trn and inf servers that use them. This position is for a Software Engineer that will lead the development of various services that will aid in optimization, analysis and release of machine learning workloads and artifacts. This candidate must have had experience leading distributed systems and machine learning related projects, preferably starting from architecture through several generations of delivery to customers. Deep knowledge of optimization, resource management, scheduling are needed. The ideal candidate will have experience working on services like EC2, EKS, Lambda in AWS or similar services on other cloud providers. Key job responsibilities This engineer will lead the design and implementation of new tools, pipelines and automation, will work with developers, system architects, hardware engineers and users both within and external to Amazon to ensure compatibility of this new toolset with existing and next-generation AI accelerators. Design, implement, and maintain CI/CD pipelines to automate the software release process. Collaborate with development teams to integrate new software releases. Infrastructure Management: Manage and automate infrastructure provisioning. Ensure high availability and scalability of systems through effective infrastructure management. Monitoring and Optimization: Implement monitoring solutions to track system performance. Identify bottlenecks and optimize system performance. Security and Compliance: Implement security best practices in the DevOps pipeline. Conduct regular vulnerability assessments and risk management. A day in the life As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also: Build high-impact solutions to deliver to our large customer base. Participate in design discussions, code review, and communicate with internal and external stakeholders. Work cross-functionally to help drive business decisions with your technical input. Work in a startup-like development environment, where you're always working on the most important stuff. About the team The Neuron Infra Services team fosters a builder's culture where experimentation is encouraged, and impact is measurable. We emphasize collaboration, technical ownership, and continuous learning. Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Knowledge of system performance, memory management, and parallel computing principles - Experience in debugging, profiling, and implementing software engineering best practices in large-scale systems, or experience debugging, profiling, and implementing best software engineering practices in large-scale systems - Experience with AWS or cloud technologies PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - Knowledge of fundamentals of networking, security, databases (relational or NoSQL), operating systems (Unix, Linux, and/or Windows) - Fundamentals of Machine learning and LLMs, their architecture along with work experience on certain LLM models. Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
Software Development Engineer - Silicon Development Infrastructure
Annapurna Labs (U.S.) Inc. Austin, Texas
We're seeking a Software Development Engineer to help architect, build, and operate the infrastructure that accelerates silicon development at Annapurna Labs. In this role, you'll contribute to the platforms, tooling, and automation that enable our chip design teams to iterate faster, validate more thoroughly, and bring transformative silicon to market. You'll work at the intersection of cloud infrastructure, high-performance computing, and electronic design automation-building systems that directly impact AWS's ability to innovate in custom silicon. This is a unique opportunity to grow your skills in infrastructure that supports chip development while working with world-class engineers across hardware, software, and operations disciplines. Key job responsibilities Customer-Focused Infrastructure Development • Partner with silicon design, verification, emulation, and software teams to understand their development workflows, pain points, and iteration cycles. • Build tooling and automation that eliminates manual toil and reduces time-to-results. • Gather continuous feedback from internal customers and rapidly iterate on solutions. Benchmark infrastructure based on silicon development workflows to provide internal customers with the optimal resources for silicon development. Own Platform Delivery and Operations • Design, implement, and operate cloud infrastructure and high-performance computing clusters using schedulers like Slurm. • Build and maintain CI/CD pipelines for infrastructure-as-code and service deployments with comprehensive testing and safe rollback mechanisms. • Take ownership of platform reliability, performance, and cost efficiency from initial design through production operation. Drive Results Through Automation and Observability • Develop monitoring, diagnostics, and alerting systems that surface actionable insights on efficiency, utilization, reliability, and cost trends. • Establish incident response processes, runbooks, and documentation that enable operational excellence. • Proactively anticipate system failures and implement preventive measures, reducing operational toil and improving system resilience. A day in the life Each day you will work with some of the best engineers in the industry to develop Machine Learning Accelerators. On-site in Austin, Texas, you will be part of the team that develops custom silicon and contribute to the infrastructure that enables this innovation. You might start your day investigating anomalies in job completion rates or resource utilization patterns. You could spend your morning collaborating with a design verification team to optimize their regression workflows, identifying bottlenecks and proposing improvements. In the afternoon, you might be building new tooling that simplifies infrastructure access for emulation teams, or contributing to monitoring dashboards that give teams real-time visibility into their development velocity. You'll participate in design reviews, contribute to postmortems when incidents occur, and continuously refine the systems that accelerate the path from RTL to silicon. Throughout the day, you'll balance immediate customer needs-unblocking a team waiting for compute capacity -with longer-term platform investments. You'll write code, review infrastructure-as-code changes, and collaborate across teams who depend on the systems you build. Take a look inside our labs to see what you will learn at Annapurna Labs: • custom-chips • About the team At Annapurna Labs, your infrastructure work directly enables breakthrough innovations in custom silicon that power AWS and transform industries. You'll collaborate with world-class chip designers, verification engineers, and software developers who are pushing the boundaries of what's possible. We offer the resources and scale of AWS with the innovation culture and technical depth of a focused silicon team. If you're passionate about building infrastructure that accelerates innovation, thrive on customer obsession and ownership, and want to see your work enable the next generation of AWS silicon -we want to hear from you BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - 2+ years of non-internship professional software development experience - 2+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience - 3+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering experience - Knowledge of systems engineering fundamentals (networking, storage, operating systems) - Experience programming with at least one modern language such as C++, C#, Java, - Python, Golang, PowerShell, Ruby - Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust with demonstrated ability to write production-quality, maintainable code - • Experience utilizing AWS cloud solutions in a DevOps environment with infrastructure as code (CloudFormation, Terraform, CDK) - • Experience with Linux/Unix - • Experience in automating, deploying, and supporting large-scale infrastructure - • Experience with high-performance computing (HPC) clusters using workload schedulers like Slurm - • Familiarity with semiconductor development workflows or electronic design automation (EDA) environments - • Experience building services using AWS products - • Experience with CI/CD pipelines and build processes - • Experience with monitoring, observability, and incident management at scale Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, TX, Austin - 143 400.00 USD annually
09/20/2026
Full time
We're seeking a Software Development Engineer to help architect, build, and operate the infrastructure that accelerates silicon development at Annapurna Labs. In this role, you'll contribute to the platforms, tooling, and automation that enable our chip design teams to iterate faster, validate more thoroughly, and bring transformative silicon to market. You'll work at the intersection of cloud infrastructure, high-performance computing, and electronic design automation-building systems that directly impact AWS's ability to innovate in custom silicon. This is a unique opportunity to grow your skills in infrastructure that supports chip development while working with world-class engineers across hardware, software, and operations disciplines. Key job responsibilities Customer-Focused Infrastructure Development • Partner with silicon design, verification, emulation, and software teams to understand their development workflows, pain points, and iteration cycles. • Build tooling and automation that eliminates manual toil and reduces time-to-results. • Gather continuous feedback from internal customers and rapidly iterate on solutions. Benchmark infrastructure based on silicon development workflows to provide internal customers with the optimal resources for silicon development. Own Platform Delivery and Operations • Design, implement, and operate cloud infrastructure and high-performance computing clusters using schedulers like Slurm. • Build and maintain CI/CD pipelines for infrastructure-as-code and service deployments with comprehensive testing and safe rollback mechanisms. • Take ownership of platform reliability, performance, and cost efficiency from initial design through production operation. Drive Results Through Automation and Observability • Develop monitoring, diagnostics, and alerting systems that surface actionable insights on efficiency, utilization, reliability, and cost trends. • Establish incident response processes, runbooks, and documentation that enable operational excellence. • Proactively anticipate system failures and implement preventive measures, reducing operational toil and improving system resilience. A day in the life Each day you will work with some of the best engineers in the industry to develop Machine Learning Accelerators. On-site in Austin, Texas, you will be part of the team that develops custom silicon and contribute to the infrastructure that enables this innovation. You might start your day investigating anomalies in job completion rates or resource utilization patterns. You could spend your morning collaborating with a design verification team to optimize their regression workflows, identifying bottlenecks and proposing improvements. In the afternoon, you might be building new tooling that simplifies infrastructure access for emulation teams, or contributing to monitoring dashboards that give teams real-time visibility into their development velocity. You'll participate in design reviews, contribute to postmortems when incidents occur, and continuously refine the systems that accelerate the path from RTL to silicon. Throughout the day, you'll balance immediate customer needs-unblocking a team waiting for compute capacity -with longer-term platform investments. You'll write code, review infrastructure-as-code changes, and collaborate across teams who depend on the systems you build. Take a look inside our labs to see what you will learn at Annapurna Labs: • custom-chips • About the team At Annapurna Labs, your infrastructure work directly enables breakthrough innovations in custom silicon that power AWS and transform industries. You'll collaborate with world-class chip designers, verification engineers, and software developers who are pushing the boundaries of what's possible. We offer the resources and scale of AWS with the innovation culture and technical depth of a focused silicon team. If you're passionate about building infrastructure that accelerates innovation, thrive on customer obsession and ownership, and want to see your work enable the next generation of AWS silicon -we want to hear from you BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - 2+ years of non-internship professional software development experience - 2+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience - 3+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering experience - Knowledge of systems engineering fundamentals (networking, storage, operating systems) - Experience programming with at least one modern language such as C++, C#, Java, - Python, Golang, PowerShell, Ruby - Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust with demonstrated ability to write production-quality, maintainable code - • Experience utilizing AWS cloud solutions in a DevOps environment with infrastructure as code (CloudFormation, Terraform, CDK) - • Experience with Linux/Unix - • Experience in automating, deploying, and supporting large-scale infrastructure - • Experience with high-performance computing (HPC) clusters using workload schedulers like Slurm - • Familiarity with semiconductor development workflows or electronic design automation (EDA) environments - • Experience building services using AWS products - • Experience with CI/CD pipelines and build processes - • Experience with monitoring, observability, and incident management at scale Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, TX, Austin - 143 400.00 USD annually
AI & Emerging Tech Specialist, Emerging Technology
Amazon Web Services, Inc. Herndon, Virginia
Amazon Web Services (AWS) is the leading cloud provider. AWS runs a globally distributed and resilient environment, operating at massive scale and enabling businesses and government agencies to run their operations and applications on AWS's multi-tenant infrastructure. Through AWS' virtual infrastructure, customers all over the world innovate faster with emerging technologies like AI/machine learning including generative AI, high performance computing, quantum, and more. Our World Wide Public Sector (WWPS) team is looking for an exceptional sales candidate who is excited about accelerating customers' journey to transform their mission areas and business functions with emerging technology. The candidate will join our fast growing AWS National Security Emerging Technology team to support sensitive and critical US Defense and Intelligence cloud technologies and activities. In this AI & Emerging Tech Specialist role, you will focus on developing customer use cases, driving adoption, and shaping the go-to-market strategy across multiple customer agencies in close collaboration with each account team. You will work backwards from customers' mission needs to develop integrated solutions that leverage AWS and partner data, analytics, generative AI, and emerging technology capabilities. As an AI & Emerging Tech Specialist, you will partner with customer executives and builders as they adopt DataOps, MLOps, and AI governance best practices to operationalize AI at scale in secure and classified environments. This position requires that the candidate selected be a US Citizen and must currently possess an active Top Secret security clearance. The position further requires that, after start, the selected candidate obtain and maintain an active TS/SCI security clearance with polygraph and satisfy other security related requirements. Key job responsibilities - Prospect into new AI and emerging technology customers and buying centers by working closely with AWS account teams and partners to identify mission-driven use cases for generative AI, agentic AI, ML at the edge, quantum, and other emerging capabilities. - Serve as a subject matter expert in AI and emerging technologies - including generative AI, foundation models, and associated AWS services - providing specialized sales guidance, technical enablement, and thought leadership to account teams and customers. - Meet regularly with national security executives and mission owners to understand their mission requirements, existing architectures, and technology roadmaps, and introduce AWS AI, data, analytics, and emerging technology services that accelerate mission outcomes. - Negotiate and close large, complex AI and emerging technology opportunities, working across multiple agencies and buying centers to drive adoption at scale. - Work closely with Amazon business leaders - including customer account teams, hardware and software engineering leaders, professional services, and technical program managers - to align emerging technology solutions with customer mission needs. - Represent the voice of the customer; collaborate with field and central teams to bring customer feedback to product teams. Lead curation of custom feature requests, regional availability requirements, and unique mission use cases across classified and unclassified environments. - Identify opportunities for Marketing and Public Policy engagements that position AWS as a trusted partner for AI and emerging technologies in the national security community. - Leverage Defense and Professional Associations to establish AWS as a thought and industry leader in AI and emerging technologies, deepening customer intimacy, nurturing existing partner relationships, developing new partnerships, and uncovering new opportunities. A day in the life This is a highly dynamic role, and as such, your typical day could include 2-3 morning meetings with agency executives, program officers, and builders. In the afternoon, you might be leading an internal strategy or working session with key stakeholders from NatSec account teams, service teams, or Professional Services. About the team The team you would join specializes in AI and Emerging Technologies for National Security, with a core focus on generative AI, frontier AI, ML at the edge, quantum, mission networking, and advanced computing. We pride ourselves on our deep technical expertise and mission knowledge. Members of our team dive deep into their technical specialties, stay ahead of trends in mission adoption of emerging technologies, and develop best practices for accelerating responsible adoption across classified and unclassified environments. As a high-performing sales organization embedded within the NatSec community, the team serves as a trusted partner to end users, agency leadership, and the policy community - helping customers navigate the rapidly evolving AI and emerging technology landscape to deliver mission impact. Diverse Experiences Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Why AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, inspire us to never stop embracing our uniqueness. Mentorship and Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - Bachelor's degree in a relevant field or equivalent work experience - 3+ years of technology platform sales with an understanding of government IT, data centers, cloud services and cloud adoption experience - Experience working cross-functionally with tech and non-tech teams - Current, active US Government Security Clearance of Top Secret or above PREFERRED QUALIFICATIONS - 5+ years of direct sales or business development in software, cloud or SaaS markets selling to C-level executives experience - Experience with one or more of the following domains: analytics, security, storage, DevOps, application development, or machine learning - Experience selling AI/ML solutions - Experience interpreting data and making business recommendations - Experience developing, negotiating and executing business agreements Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, MD, Jessup - 92 000.00 USD annually USA, VA, Arlington - 92 000.00 USD annually USA, VA, Herndon - 92 000.00 USD annually
09/20/2026
Full time
Amazon Web Services (AWS) is the leading cloud provider. AWS runs a globally distributed and resilient environment, operating at massive scale and enabling businesses and government agencies to run their operations and applications on AWS's multi-tenant infrastructure. Through AWS' virtual infrastructure, customers all over the world innovate faster with emerging technologies like AI/machine learning including generative AI, high performance computing, quantum, and more. Our World Wide Public Sector (WWPS) team is looking for an exceptional sales candidate who is excited about accelerating customers' journey to transform their mission areas and business functions with emerging technology. The candidate will join our fast growing AWS National Security Emerging Technology team to support sensitive and critical US Defense and Intelligence cloud technologies and activities. In this AI & Emerging Tech Specialist role, you will focus on developing customer use cases, driving adoption, and shaping the go-to-market strategy across multiple customer agencies in close collaboration with each account team. You will work backwards from customers' mission needs to develop integrated solutions that leverage AWS and partner data, analytics, generative AI, and emerging technology capabilities. As an AI & Emerging Tech Specialist, you will partner with customer executives and builders as they adopt DataOps, MLOps, and AI governance best practices to operationalize AI at scale in secure and classified environments. This position requires that the candidate selected be a US Citizen and must currently possess an active Top Secret security clearance. The position further requires that, after start, the selected candidate obtain and maintain an active TS/SCI security clearance with polygraph and satisfy other security related requirements. Key job responsibilities - Prospect into new AI and emerging technology customers and buying centers by working closely with AWS account teams and partners to identify mission-driven use cases for generative AI, agentic AI, ML at the edge, quantum, and other emerging capabilities. - Serve as a subject matter expert in AI and emerging technologies - including generative AI, foundation models, and associated AWS services - providing specialized sales guidance, technical enablement, and thought leadership to account teams and customers. - Meet regularly with national security executives and mission owners to understand their mission requirements, existing architectures, and technology roadmaps, and introduce AWS AI, data, analytics, and emerging technology services that accelerate mission outcomes. - Negotiate and close large, complex AI and emerging technology opportunities, working across multiple agencies and buying centers to drive adoption at scale. - Work closely with Amazon business leaders - including customer account teams, hardware and software engineering leaders, professional services, and technical program managers - to align emerging technology solutions with customer mission needs. - Represent the voice of the customer; collaborate with field and central teams to bring customer feedback to product teams. Lead curation of custom feature requests, regional availability requirements, and unique mission use cases across classified and unclassified environments. - Identify opportunities for Marketing and Public Policy engagements that position AWS as a trusted partner for AI and emerging technologies in the national security community. - Leverage Defense and Professional Associations to establish AWS as a thought and industry leader in AI and emerging technologies, deepening customer intimacy, nurturing existing partner relationships, developing new partnerships, and uncovering new opportunities. A day in the life This is a highly dynamic role, and as such, your typical day could include 2-3 morning meetings with agency executives, program officers, and builders. In the afternoon, you might be leading an internal strategy or working session with key stakeholders from NatSec account teams, service teams, or Professional Services. About the team The team you would join specializes in AI and Emerging Technologies for National Security, with a core focus on generative AI, frontier AI, ML at the edge, quantum, mission networking, and advanced computing. We pride ourselves on our deep technical expertise and mission knowledge. Members of our team dive deep into their technical specialties, stay ahead of trends in mission adoption of emerging technologies, and develop best practices for accelerating responsible adoption across classified and unclassified environments. As a high-performing sales organization embedded within the NatSec community, the team serves as a trusted partner to end users, agency leadership, and the policy community - helping customers navigate the rapidly evolving AI and emerging technology landscape to deliver mission impact. Diverse Experiences Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Why AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, inspire us to never stop embracing our uniqueness. Mentorship and Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - Bachelor's degree in a relevant field or equivalent work experience - 3+ years of technology platform sales with an understanding of government IT, data centers, cloud services and cloud adoption experience - Experience working cross-functionally with tech and non-tech teams - Current, active US Government Security Clearance of Top Secret or above PREFERRED QUALIFICATIONS - 5+ years of direct sales or business development in software, cloud or SaaS markets selling to C-level executives experience - Experience with one or more of the following domains: analytics, security, storage, DevOps, application development, or machine learning - Experience selling AI/ML solutions - Experience interpreting data and making business recommendations - Experience developing, negotiating and executing business agreements Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, MD, Jessup - 92 000.00 USD annually USA, VA, Arlington - 92 000.00 USD annually USA, VA, Herndon - 92 000.00 USD annually
Mastercard
Senior Site Reliability Engineer
Mastercard O Fallon, Missouri
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team - one that makes better decisions, drives innovation, and delivers better business results. Technology at Mastercard What we create today will define tomorrow. Revolutionary technologies that reshape the digital economy to be more connected and inclusive than ever before. Safer, faster, more sustainable. And we need the best people to do it. Technologists who are energized by the challenges of a truly global network. With the talent and vision to create the critical systems and products that power global commerce and connect people everywhere to the vital goods and services they need every day. Working at Mastercard means being part of a unique culture. Inclusive and diverse, a rich collaboration of ideas and perspectives. A place that celebrates your strengths, values your experiences, and offers you the flexibility to shape a career across disciplines and continents. And the opportunity to work alongside experts and leaders at every level of the business, improving what exists, and inventing what's next. About the Role The Business Operations team is seeking a highly motivated and experienced Senior Site Reliability Engineer (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor. The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications. Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle. As part of the Business Operations team, you will: • Independently execute key elements of projects/processes within the Site Reliability Engineering area by applying in-depth knowledge of their discipline and area best practices to effectively resolve problems and roadblocks as they occur. • Assist in evaluating operational requirements and developing technical solutions within existing frameworks. • Support automation and scripting efforts to improve operational workflows and incident response processes. • Troubleshoot and resolve routine and some complex system issues, escalating when necessary to maintain system health. • Contribute to documentation, knowledge sharing, and best practices to enhance team operational procedures. • Collaborate with development teams and stakeholders to ensure reliability solutions align with technical and business needs. • Participate in reviews and quality assurance activities to uphold system stability standards. • May contribute to solution development for new products/services and/or manage smaller project/initiatives as an experienced individual contributor with specialized knowledge within the Site Reliability Engineering area. Role qualifications: The ideal candidate will apply the following skills independently in routine and moderately complex situations, requiring occasional guidance typically only in unfamiliar or highly complex scenarios. They will demonstrate growing consistency and reliability in applying the skills. • Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement. • Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar. • Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability. • Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency. • Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand • DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations. • Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations. • Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage. • IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs. • Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability. Mastercard is a merit-based, inclusive, equal opportunity employer that considers applicants without regard to gender, gender identity, sexual orientation, race, ethnicity, disabled or veteran status, or any other characteristic protected by law. We hire the most qualified candidate for the role. In the US or Canada, if you require accommodations or assistance to complete the online application process or during the recruitment process, please contact and identify the type of accommodation or assistance you are requesting. Do not include any medical or health information in this email. The Reasonable Accommodations team will respond to your email promptly. Corporate Security Responsibility All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must: Abide by Mastercard's security policies and practices; Ensure the confidentiality and integrity of the information being accessed; Report any suspected information security violation or breach, and Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines. In line with Mastercard's total compensation philosophy and assuming that the job will be performed in the US, the successful candidate will be offered a competitive base salary and may be eligible for an annual bonus or commissions depending on the role . click apply for full job details
09/19/2026
Full time
Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits everyone, everywhere, by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships, and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. We cultivate a culture of inclusion for all employees that respects their individual strengths, views, and experiences. We believe that our differences enable us to be a better team - one that makes better decisions, drives innovation, and delivers better business results. Technology at Mastercard What we create today will define tomorrow. Revolutionary technologies that reshape the digital economy to be more connected and inclusive than ever before. Safer, faster, more sustainable. And we need the best people to do it. Technologists who are energized by the challenges of a truly global network. With the talent and vision to create the critical systems and products that power global commerce and connect people everywhere to the vital goods and services they need every day. Working at Mastercard means being part of a unique culture. Inclusive and diverse, a rich collaboration of ideas and perspectives. A place that celebrates your strengths, values your experiences, and offers you the flexibility to shape a career across disciplines and continents. And the opportunity to work alongside experts and leaders at every level of the business, improving what exists, and inventing what's next. About the Role The Business Operations team is seeking a highly motivated and experienced Senior Site Reliability Engineer (SRE) to join our team. You will play a critical role in ensuring the reliability, scalability, and performance of our applications, supporting essential services that power Mastercard's global operations. As a thought leader in your field, you will bring technical expertise, a passion for automation, and the ability to mentor. The role of the Business Operations Site Reliability Engineer is to be the production readiness steward for Mastercard products. As Business Operations SRE, we are responsible for ensuring that our platform is stable and healthy. We break down barriers to running our products by fostering developer run ownership and empowering developers to build resilient products. We support our developers during the application build phase in software run principles that include operational design, automation, capacity planning, and monitoring that leads to fault-tolerant, scalable products. We see the big picture and help create and enforce operations standards while facilitating an agile and learning culture. We support daily operations with a hyper focus on triage, root cause by understanding the business impact of our products and subsequently performing blameless post-mortems. The goal of every Business Operations team is to engage early in the development lifecycle to be more proactive and upfront in the development process, and to proactively manage production and change activities to maximize customer experience and increase the overall value of supported applications. Business Operations teams also focus on risk management by tying all our activities together with an overarching responsibility for compliance and risk mitigation across all our environments. Ultimately, the role of Business Operations is to align Product and Customer Focused priorities with Operational needs by providing continuous feedback throughout the lifecycle. As part of the Business Operations team, you will: • Independently execute key elements of projects/processes within the Site Reliability Engineering area by applying in-depth knowledge of their discipline and area best practices to effectively resolve problems and roadblocks as they occur. • Assist in evaluating operational requirements and developing technical solutions within existing frameworks. • Support automation and scripting efforts to improve operational workflows and incident response processes. • Troubleshoot and resolve routine and some complex system issues, escalating when necessary to maintain system health. • Contribute to documentation, knowledge sharing, and best practices to enhance team operational procedures. • Collaborate with development teams and stakeholders to ensure reliability solutions align with technical and business needs. • Participate in reviews and quality assurance activities to uphold system stability standards. • May contribute to solution development for new products/services and/or manage smaller project/initiatives as an experienced individual contributor with specialized knowledge within the Site Reliability Engineering area. Role qualifications: The ideal candidate will apply the following skills independently in routine and moderately complex situations, requiring occasional guidance typically only in unfamiliar or highly complex scenarios. They will demonstrate growing consistency and reliability in applying the skills. • Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement. • Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar. • Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability. • Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency. • Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand • DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations. • Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations. • Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage. • IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs. • Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability. Mastercard is a merit-based, inclusive, equal opportunity employer that considers applicants without regard to gender, gender identity, sexual orientation, race, ethnicity, disabled or veteran status, or any other characteristic protected by law. We hire the most qualified candidate for the role. In the US or Canada, if you require accommodations or assistance to complete the online application process or during the recruitment process, please contact and identify the type of accommodation or assistance you are requesting. Do not include any medical or health information in this email. The Reasonable Accommodations team will respond to your email promptly. Corporate Security Responsibility All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must: Abide by Mastercard's security policies and practices; Ensure the confidentiality and integrity of the information being accessed; Report any suspected information security violation or breach, and Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines. In line with Mastercard's total compensation philosophy and assuming that the job will be performed in the US, the successful candidate will be offered a competitive base salary and may be eligible for an annual bonus or commissions depending on the role . click apply for full job details
Principal Machine Learning Engineer
Vail Resorts Broomfield, Colorado
Job Description Job Description Our mission is to create the Experience of a Lifetime for our employees, so they can, in turn, create the Experience of a Lifetime for our guests. We own and operate the most renowned destination resorts in the world as well as regional and local ski areas outside major cities, and connect them all through one unrivaled network. We are looking for ambitious leaders, innovators and creators to join our talented team. If you're ready to pursue your fullest potential, we want to get to know you! Candidates for year-round positions are reviewed on a rolling basis. Applications will be accepted up to 90 days after the posting date, or until the position is filled (whichever is first). Job Summary: We are looking for a curious, driven, innovative machine learning engineer who takes initiative to solve problems and create environments that accelerate the development, deployment, and usage of data science models and AI to drive greater organizational impact. The Data Science & Data Engineering team within the Enterprise Analytics organization builds data assets, predictive models, analytical applications, and platforms across the organization. Our team collaborates with business stakeholders, analysts, and technology teams to tackle high-impact use cases with state-of-the-art models and tools to grow the business, streamline costs, and improve guest experiences. Job Specifications: Starting Wage: $140,000 - $185,000 + Annual Bonus Employment Type: Year Round Shift Type: Full Time hours Minimum Age: At least 18 years of age Housing Availability: No Job Responsibilities: Productionize ML models developed by data science into reliable, monitored, maintainable systems. Build model data foundations that ensure training, inference, monitoring, and analytics data are trustworthy and scalable. Architect ML platform patterns in Databricks that bring reliability, consistency, governance, performance, and cost discipline to ML and data workflows. Identify and scope opportunities for ML engineering across the business for high-impact. Develop reusable tools , libraries, standards, documentation, and production-readiness practices to enable data science and data engineering teams. Develop analytical and model-powered applications that turn data and ML outputs into usable business workflows for end users. Prepare the platform for future AI engineering , including LLM and agent-based systems, as the organization matures. Provide technical leadership and mentoring across engineering, architecture, and development including design and code reviews. Job Requirements: Technical Skills: Quantitative Foundation : B.S. degree in a quantitative field (e.g., Computer Science, Mathematics, Statistics, Economics, Operations Research, Engineering). Software Engineering Fundamentals: write clean, modular, testable, maintainable code and understand how to structure production-grade systems rather than one-off notebooks or scripts. Python and SQL Proficiency: strong in Python and SQL for building data pipelines, automation, model integrations, analytical workflows, and production services. Data Modeling and Pipeline Design: understand how to design reliable, well-structured data assets, including curated tables, feature datasets, batch pipelines, orchestration, data quality checks, and lineage. ML Lifecycle Fluency: understand the full model lifecycle: data collection, exploration, model development, validation, deployment, monitoring, retraining, and retirement. Production ML Patterns: understand core MLOps patterns such as model registries, feature/data versioning, reproducible environments, testing/validation, monitoring, and rollback. Cloud and Platform Engineering: You are comfortable working in cloud-based data and ML environments and understand the foundations of permissions, environments, jobs, services, storage, networking, and cost-aware architecture. Databricks Expertise : You're familiar and experienced with the core parts of Spark, Unity Catalog, Delta Lake, Databricks Workflows, MLflow, model registry patterns, job/cluster optimization, and governance. DevOps Practices: You use modern engineering practices such as Git, CI/CD, automated testing, code review, dependency management, environment management, and observability. Application Development : You can build applications, APIs, dashboards, or workflow tools that sit on top of data and model outputs. System Design: You can reason through tradeoffs across reliability, latency, scale, cost, governance, maintainability, and ease of use. Soft Skills: Curious : bring intellectual curiosity, an inquisitive nature, and a desire to deepen your knowledge and continue learning. Ownership : take responsibility to proactively advance projects, contribute to the organization, and develop the best solutions. Communication: explain technical concepts, risks, tradeoffs, and recommendations clearly to technical and non-technical audiences. Collaboration: work effectively cross-functionally with data scientists, data engineers, analysts, application engineers, product partners, and business stakeholders. Pragmatism : You know how to balance ideal architecture with business urgency, team maturity, operational constraints, and the need to ship. Preferred qualifications: A graduate degree (Masters or PhD) in a quantitative field Experience with dbt (Core) for modular data modeling, including testing, documentation, and dependency management Experience with AI engineer to use, build, and monitor agentic solutions The expected Total Compensation for this role is $140,000 - $185,000 + Annual Bonus. Individual compensation decisions are based on a variety of factors. Job Benefits Ski/Mountain Perks! Free passes for employees, employee discounted lift tickets for friends and family AND free ski lessons MORE employee discounts on lodging, food, gear, and mountain shuttles 401(k) Retirement Plan Employee Assistance Program Excellent training and professional development Full Time roles are eligible for the above, plus: Health Insurance; Medical Insurance, Dental Insurance, and Vision Insurance plans (for eligible seasonal employees after working 500 hours) Free ski passes for dependents Critical Illness and Accident plans Employees can work remotely from British Columbia, Washington D.C., and the 16 U.S. states in which we currently operate. This includes: California, Colorado, Indiana, Michigan, Minnesota, Missouri, New Hampshire, New York, Nevada, Ohio, Pennsylvania, Utah, Vermont, Washington State, Wisconsin, and Wyoming. Please note that the ability to work in person or off-site, and the particulars related to such work, are subject to change at any time; and, accordingly, the Company reserves the right to change its policies and/or require in-person/in-office work or off-site work at any time in its sole discretion. In completing this application, and when submitting related documentation, applicants may redact information that identifies their age, date of birth, and/or dates of attendance at or graduation from an educational institution. We follow all federal, state, and local laws including restrictions on child/minor labor. Minors hired into this position will not be asked or permitted to engage in any activities restricted to adult workers. Vail Resorts is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability, protected veteran status or any other status protected by applicable law. Requisition ID 517322 Reference Date: 09/05/2026 Job Code Function: Data Science
09/19/2026
Full time
Job Description Job Description Our mission is to create the Experience of a Lifetime for our employees, so they can, in turn, create the Experience of a Lifetime for our guests. We own and operate the most renowned destination resorts in the world as well as regional and local ski areas outside major cities, and connect them all through one unrivaled network. We are looking for ambitious leaders, innovators and creators to join our talented team. If you're ready to pursue your fullest potential, we want to get to know you! Candidates for year-round positions are reviewed on a rolling basis. Applications will be accepted up to 90 days after the posting date, or until the position is filled (whichever is first). Job Summary: We are looking for a curious, driven, innovative machine learning engineer who takes initiative to solve problems and create environments that accelerate the development, deployment, and usage of data science models and AI to drive greater organizational impact. The Data Science & Data Engineering team within the Enterprise Analytics organization builds data assets, predictive models, analytical applications, and platforms across the organization. Our team collaborates with business stakeholders, analysts, and technology teams to tackle high-impact use cases with state-of-the-art models and tools to grow the business, streamline costs, and improve guest experiences. Job Specifications: Starting Wage: $140,000 - $185,000 + Annual Bonus Employment Type: Year Round Shift Type: Full Time hours Minimum Age: At least 18 years of age Housing Availability: No Job Responsibilities: Productionize ML models developed by data science into reliable, monitored, maintainable systems. Build model data foundations that ensure training, inference, monitoring, and analytics data are trustworthy and scalable. Architect ML platform patterns in Databricks that bring reliability, consistency, governance, performance, and cost discipline to ML and data workflows. Identify and scope opportunities for ML engineering across the business for high-impact. Develop reusable tools , libraries, standards, documentation, and production-readiness practices to enable data science and data engineering teams. Develop analytical and model-powered applications that turn data and ML outputs into usable business workflows for end users. Prepare the platform for future AI engineering , including LLM and agent-based systems, as the organization matures. Provide technical leadership and mentoring across engineering, architecture, and development including design and code reviews. Job Requirements: Technical Skills: Quantitative Foundation : B.S. degree in a quantitative field (e.g., Computer Science, Mathematics, Statistics, Economics, Operations Research, Engineering). Software Engineering Fundamentals: write clean, modular, testable, maintainable code and understand how to structure production-grade systems rather than one-off notebooks or scripts. Python and SQL Proficiency: strong in Python and SQL for building data pipelines, automation, model integrations, analytical workflows, and production services. Data Modeling and Pipeline Design: understand how to design reliable, well-structured data assets, including curated tables, feature datasets, batch pipelines, orchestration, data quality checks, and lineage. ML Lifecycle Fluency: understand the full model lifecycle: data collection, exploration, model development, validation, deployment, monitoring, retraining, and retirement. Production ML Patterns: understand core MLOps patterns such as model registries, feature/data versioning, reproducible environments, testing/validation, monitoring, and rollback. Cloud and Platform Engineering: You are comfortable working in cloud-based data and ML environments and understand the foundations of permissions, environments, jobs, services, storage, networking, and cost-aware architecture. Databricks Expertise : You're familiar and experienced with the core parts of Spark, Unity Catalog, Delta Lake, Databricks Workflows, MLflow, model registry patterns, job/cluster optimization, and governance. DevOps Practices: You use modern engineering practices such as Git, CI/CD, automated testing, code review, dependency management, environment management, and observability. Application Development : You can build applications, APIs, dashboards, or workflow tools that sit on top of data and model outputs. System Design: You can reason through tradeoffs across reliability, latency, scale, cost, governance, maintainability, and ease of use. Soft Skills: Curious : bring intellectual curiosity, an inquisitive nature, and a desire to deepen your knowledge and continue learning. Ownership : take responsibility to proactively advance projects, contribute to the organization, and develop the best solutions. Communication: explain technical concepts, risks, tradeoffs, and recommendations clearly to technical and non-technical audiences. Collaboration: work effectively cross-functionally with data scientists, data engineers, analysts, application engineers, product partners, and business stakeholders. Pragmatism : You know how to balance ideal architecture with business urgency, team maturity, operational constraints, and the need to ship. Preferred qualifications: A graduate degree (Masters or PhD) in a quantitative field Experience with dbt (Core) for modular data modeling, including testing, documentation, and dependency management Experience with AI engineer to use, build, and monitor agentic solutions The expected Total Compensation for this role is $140,000 - $185,000 + Annual Bonus. Individual compensation decisions are based on a variety of factors. Job Benefits Ski/Mountain Perks! Free passes for employees, employee discounted lift tickets for friends and family AND free ski lessons MORE employee discounts on lodging, food, gear, and mountain shuttles 401(k) Retirement Plan Employee Assistance Program Excellent training and professional development Full Time roles are eligible for the above, plus: Health Insurance; Medical Insurance, Dental Insurance, and Vision Insurance plans (for eligible seasonal employees after working 500 hours) Free ski passes for dependents Critical Illness and Accident plans Employees can work remotely from British Columbia, Washington D.C., and the 16 U.S. states in which we currently operate. This includes: California, Colorado, Indiana, Michigan, Minnesota, Missouri, New Hampshire, New York, Nevada, Ohio, Pennsylvania, Utah, Vermont, Washington State, Wisconsin, and Wyoming. Please note that the ability to work in person or off-site, and the particulars related to such work, are subject to change at any time; and, accordingly, the Company reserves the right to change its policies and/or require in-person/in-office work or off-site work at any time in its sole discretion. In completing this application, and when submitting related documentation, applicants may redact information that identifies their age, date of birth, and/or dates of attendance at or graduation from an educational institution. We follow all federal, state, and local laws including restrictions on child/minor labor. Minors hired into this position will not be asked or permitted to engage in any activities restricted to adult workers. Vail Resorts is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability, protected veteran status or any other status protected by applicable law. Requisition ID 517322 Reference Date: 09/05/2026 Job Code Function: Data Science
Solutions Architect - EA (Azure)
Code Plus Inc Huntsville, Alabama
Job Description Job Description Why CODEplus CODEplus is a software engineering and modernization company with over 31 years supporting federal agencies including DoD, NASA, DOE, NRC, and USPS. We combine the agility and innovation of a small business with mature engineering, Agile, and DevSecOps delivery practices to support mission-critical modernization efforts. Position Title: Solutions Architect - EA (Azure) Various position levels for each position: Junior-level - 2-3 years of experience Mid-level - 4-8 years of experience Senior-level - 10+ years of experience Location: Huntsville, AL (Hybrid/On-Site preferred, possibly remote) Clearance: Active Secret Clearance Required Position Overview CODEplus is seeking a Solutions Architect - EA (Azure) to design, lead, and execute large-scale cloud modernization efforts for Department of Defense and federal customers. This role is ideal for a hands-on cloud architect with deep engineering expertise , who has led enterprise migrations, designed secure cloud platforms, and advised senior stakeholders on cloud strategy and transformation . You will serve as the technical authority for Azure architecture , driving solution design across complex, mission-critical environments , while actively contributing to implementation alongside Agile and DevSecOps teams. This position blends architecture leadership, hands-on engineering, and customer advisory , consistent with CODEplus's engineering-first culture. Key Responsibilities Azure Architecture & Platform Design Lead architecture and design of enterprise-scale Azure solutions , including Azure Government and hybrid environments Design and implement Azure landing zones , multi-region architectures, and high-availability solutions Define reusable cloud patterns, reference architectures, and platform standards Architect solutions aligned with Well-Architected Framework principles (Azure CAF) , performance, cost optimization, and operational excellence Ensure compliance with DoD security requirements (IL4/IL5, RMF, Zero Trust, NIST 800-53) Cloud Modernization & Migration Lead large-scale cloud migrations and transformations , including legacy and cloud-native workloads Architect modernization approaches including: Rehosting, replatforming, and refactoring strategies Containerization and microservices (AKS) Serverless and event-driven architectures Support migration planning for 100+ application portfolios , including dependency mapping and sequencing Hands-on Engineering & DevSecOps Actively contribute to implementation, including: Infrastructure as Code ( Terraform, Bicep, ARM ) CI/CD pipelines and DevSecOps automation Security automation and compliance enforcement Develop or guide development of automation scripts (Python, PowerShell) for infrastructure and operations Lead design reviews, code reviews, and ensure architectural integrity across deployments Agile Leadership & Technical Execution Provide technical leadership to Agile/SAFe teams , supporting PI planning, backlog refinement, and sprint execution Mentor engineers in cloud engineering, DevOps, and architecture best practices Resolve complex system-level challenges across performance, scalability, and security Customer Advisory & Stakeholder Engagement Serve as a trusted advisor to government customers and senior leadership Translate mission needs into technical strategy and cloud architecture decisions Deliver: Executive briefings Architecture diagrams Trade studies and decision papers Participate in TIMs, design reviews, and program governance Integration & Systems Engineering Lead integration across: Cloud services On-prem / hybrid environments External mission systems Collaborate across cybersecurity, networking, software, and systems engineering teams Cost, Risk & Program Support Support cloud cost modeling, optimization strategies , and trade-off analysis Identify technical risks and drive mitigation strategies Contribute to proposals, including: Architecture approaches Basis of Estimate (BOE) Technical volumes Required Qualifications Bachelor's degree in STEM Deep, hands-on Azure architecture experience , including: Networking, identity (Entra ID), compute, storage, and security Azure Government or regulated cloud environments Demonstrated experience: Leading enterprise-scale cloud migrations or transformations Designing solutions across multi-region, high-availability environments Strong hands-on experience with: Infrastructure as Code (Terraform, Bicep, ARM) CI/CD and DevSecOps pipelines Automation scripting ( Python, PowerShell ) Proven ability to: Lead technical teams Work directly with government customers Balance technical excellence, cost, and delivery timelines Strong communication skills with experience briefing senior stakeholders Preferred Qualifications Azure certifications: Azure Solutions Architect Expert Azure Security Engineer Azure DevOps Engineer Experience with: Azure Landing Zones / CAF implementation Containers and Kubernetes ( AKS ) Serverless (Functions, Event Grid, Service Bus) RMF / NIST 800-53 / FedRAMP Platform engineering and internal developer platforms Prior experience supporting DoD, IC, or federal programs Background in systems engineering, DevOps, or software engineering Experience with large-scale migrations (100+ workloads) Experience contributing to capture, proposals, or re-competes
09/15/2026
Full time
Job Description Job Description Why CODEplus CODEplus is a software engineering and modernization company with over 31 years supporting federal agencies including DoD, NASA, DOE, NRC, and USPS. We combine the agility and innovation of a small business with mature engineering, Agile, and DevSecOps delivery practices to support mission-critical modernization efforts. Position Title: Solutions Architect - EA (Azure) Various position levels for each position: Junior-level - 2-3 years of experience Mid-level - 4-8 years of experience Senior-level - 10+ years of experience Location: Huntsville, AL (Hybrid/On-Site preferred, possibly remote) Clearance: Active Secret Clearance Required Position Overview CODEplus is seeking a Solutions Architect - EA (Azure) to design, lead, and execute large-scale cloud modernization efforts for Department of Defense and federal customers. This role is ideal for a hands-on cloud architect with deep engineering expertise , who has led enterprise migrations, designed secure cloud platforms, and advised senior stakeholders on cloud strategy and transformation . You will serve as the technical authority for Azure architecture , driving solution design across complex, mission-critical environments , while actively contributing to implementation alongside Agile and DevSecOps teams. This position blends architecture leadership, hands-on engineering, and customer advisory , consistent with CODEplus's engineering-first culture. Key Responsibilities Azure Architecture & Platform Design Lead architecture and design of enterprise-scale Azure solutions , including Azure Government and hybrid environments Design and implement Azure landing zones , multi-region architectures, and high-availability solutions Define reusable cloud patterns, reference architectures, and platform standards Architect solutions aligned with Well-Architected Framework principles (Azure CAF) , performance, cost optimization, and operational excellence Ensure compliance with DoD security requirements (IL4/IL5, RMF, Zero Trust, NIST 800-53) Cloud Modernization & Migration Lead large-scale cloud migrations and transformations , including legacy and cloud-native workloads Architect modernization approaches including: Rehosting, replatforming, and refactoring strategies Containerization and microservices (AKS) Serverless and event-driven architectures Support migration planning for 100+ application portfolios , including dependency mapping and sequencing Hands-on Engineering & DevSecOps Actively contribute to implementation, including: Infrastructure as Code ( Terraform, Bicep, ARM ) CI/CD pipelines and DevSecOps automation Security automation and compliance enforcement Develop or guide development of automation scripts (Python, PowerShell) for infrastructure and operations Lead design reviews, code reviews, and ensure architectural integrity across deployments Agile Leadership & Technical Execution Provide technical leadership to Agile/SAFe teams , supporting PI planning, backlog refinement, and sprint execution Mentor engineers in cloud engineering, DevOps, and architecture best practices Resolve complex system-level challenges across performance, scalability, and security Customer Advisory & Stakeholder Engagement Serve as a trusted advisor to government customers and senior leadership Translate mission needs into technical strategy and cloud architecture decisions Deliver: Executive briefings Architecture diagrams Trade studies and decision papers Participate in TIMs, design reviews, and program governance Integration & Systems Engineering Lead integration across: Cloud services On-prem / hybrid environments External mission systems Collaborate across cybersecurity, networking, software, and systems engineering teams Cost, Risk & Program Support Support cloud cost modeling, optimization strategies , and trade-off analysis Identify technical risks and drive mitigation strategies Contribute to proposals, including: Architecture approaches Basis of Estimate (BOE) Technical volumes Required Qualifications Bachelor's degree in STEM Deep, hands-on Azure architecture experience , including: Networking, identity (Entra ID), compute, storage, and security Azure Government or regulated cloud environments Demonstrated experience: Leading enterprise-scale cloud migrations or transformations Designing solutions across multi-region, high-availability environments Strong hands-on experience with: Infrastructure as Code (Terraform, Bicep, ARM) CI/CD and DevSecOps pipelines Automation scripting ( Python, PowerShell ) Proven ability to: Lead technical teams Work directly with government customers Balance technical excellence, cost, and delivery timelines Strong communication skills with experience briefing senior stakeholders Preferred Qualifications Azure certifications: Azure Solutions Architect Expert Azure Security Engineer Azure DevOps Engineer Experience with: Azure Landing Zones / CAF implementation Containers and Kubernetes ( AKS ) Serverless (Functions, Event Grid, Service Bus) RMF / NIST 800-53 / FedRAMP Platform engineering and internal developer platforms Prior experience supporting DoD, IC, or federal programs Background in systems engineering, DevOps, or software engineering Experience with large-scale migrations (100+ workloads) Experience contributing to capture, proposals, or re-competes
Principal Staff Engineer - Web Platform
PlanetArt Agoura Hills, California
Job Description Job Description Company and Vision PlanetArt's vision is to be the leading seller of personalized and make-on-demand products worldwide. We provide consumers with unmatched tools and content and an unparalleled end-to-end customer experience that result in high-quality, meaningful finished products and memorable celebrations of live events. The company's brands include the popular FreePrints and FreePrints Photobooks apps and the industry leading SimplytoImpress card and stationery site, as well as Personal Creations, CafePress and ISeeMe! Visit to learn more about our brands. We have more than 500 team members across multiple offices, primarily in Calabasas CA, San Diego CA, Woodridge IL, Minneapolis, MN and Pleasanton, CA. We also have team members in two company-owned offices in China, as well as in Europe. Job Overview PlanetArt is seeking a Principal Staff Engineer-Web Platform to serve as a senior technical leader within our engineering organization and a key partner to the VP of Engineering. This is a highly hands-on role focused on building, operating, and scaling high-traffic ecommerce web platforms in a fast-moving production environment. Reporting directly to the VP of Engineering, this engineer will play a critical role in architecting, developing, troubleshooting, and maintaining our LAMP-based web applications and AWS infrastructure. The ideal candidate is equally comfortable writing production code, diagnosing complex site reliability issues, managing cloud infrastructure, and leading technical problem-solving during high-severity incidents. This role requires strong operational judgment and the ability to independently own production challenges across application, database, infrastructure, and deployment layers. The engineer will collaborate closely with our China-based development organization, serving as a senior US-based technical lead responsible for cross-team coordination, code quality, architectural guidance, and production stability. This is an ideal opportunity for an experienced engineer who thrives in high-scale ecommerce environments and enjoys combining deep application engineering with modern cloud operations and production ownership. PLEASE NOTE: Candidates much be local to or willing to relocate to the Calabasas area as we operate on a hybrid work model (3 days onsite, 2 remote) What You'll Do Key Responsibilities Full-Stack Platform Engineering: Design, develop, and maintain scalable features and services across our LAMP-based ecommerce platform, with a strong focus on reliability, performance, maintainability, and operational excellence. Production Operations & Incident Response : Act as a senior technical escalation point for complex production incidents, troubleshooting issues across application, infrastructure, networking, database, CDN, and deployment layers. Lead root cause analysis and drive long-term stability improvements. AWS Infrastructure Ownership : Manage and optimize AWS infrastructure, including deployment architecture, scaling strategies, observability, security, disaster recovery, and cost efficiency. Partner closely with DevOps and engineering leadership on operational best practices. High-Scale Performance Optimization : Monitor and improve application, database, and infrastructure performance for high-traffic consumer web applications. Identify bottlenecks and implement scalable solutions to improve uptime, latency, and system resilience. Cross-Functional Technical Leadership: Partner with Product, Design, Operations, and Customer Experience teams to translate business requirements into scalable technical solutions and ensure successful project execution. Global Engineering Collaboration : Work closely with the China-based engineering team to coordinate development efforts, conduct code reviews, align on architectural direction, manage releases, and maintain strong engineering communication across time zones. Code Quality & Engineering Standards : Champion high engineering standards through code reviews, testing strategies, documentation, observability, and operational best practices. Drive continuous improvement in system reliability and development processes. Technical Mentorship & Leadership : Provide technical mentorship and architectural guidance across the engineering organization. Influence technical direction through hands-on leadership, strong execution, and collaborative problem-solving. Requirements What You Should Have Skills, Qualifications, and Requirements Senior-Level Full-Stack Engineering Experience: 5+ years of professional experience building and operating large-scale web applications, including substantial hands-on experience with the LAMP stack (Linux, Apache, MySQL, PHP). Strong AWS & Cloud Operations Expertise : Deep hands-on experience with AWS services and production cloud environments, including EC2, RDS, S3, Lambda, CloudWatch, networking, scaling, monitoring, and infrastructure troubleshooting. Ecommerce & High-Traffic Website Experience : Experience supporting high-volume consumer-facing websites or ecommerce platforms, with a strong understanding of scalability, uptime, performance optimization, and operational reliability. Production Troubleshooting Expertise : Demonstrated ability to diagnose and resolve complex production issues under pressure, including database replication issues, performance degradation, infrastructure failures, deployment issues, and site outages. Distributed Systems & Database Knowledge : Strong understanding of distributed web architectures, database performance tuning, replication strategies, caching, queuing systems, and fault-tolerant system design. Global Team Collaboration: Experience working effectively with offshore or globally distributed engineering teams, with strong communication, coordination, and cross-cultural collaboration skills. Chinese Language Skills: Ability to communicate in Mandarin (spoken or written) is highly desirable to facilitate collaboration with our China-based engineering team. Engineering Best Practices: Strong understanding of software engineering fundamentals including Git workflows, CI/CD pipelines, automated testing, observability, code review practices, and secure development standards. Ownership Mentality: Self-directed engineer with strong operational instincts, excellent judgment, and the ability to independently own critical technical initiatives from design through production support. Technical Leadership: Demonstrated ability to influence engineering direction, mentor developers, and drive technical excellence through hands-on leadership rather than direct people management. What You Can Expect Working Conditions Work is performed in an office environment with low to moderate noise levels. Position requires regular, continuous use of computer. Position requires regular sitting and standing. Position requires regular interaction with team members through the following methods: in-person, phone, Zoom, Slack, or email. May require occasional travel. This is a hybrid position; employees are expected to be in the office three days per week (Monday, Tuesday, and Thursday) with the option of working remotely two days (Wednesday and Friday). Benefits The compensation range for this position is $130,000-$220,000 annual salary. PlanetArt offers a comprehensive benefits package, including: Health, Dental, and Vision Insurance Life Insurance Pet Insurance Mental Health Insurance 401(k) with matching Comprehensive Time Off Program including Paid Time Off, Sick Days, Paid Holidays, and Floating Holidays Employee Product Discounts
09/15/2026
Full time
Job Description Job Description Company and Vision PlanetArt's vision is to be the leading seller of personalized and make-on-demand products worldwide. We provide consumers with unmatched tools and content and an unparalleled end-to-end customer experience that result in high-quality, meaningful finished products and memorable celebrations of live events. The company's brands include the popular FreePrints and FreePrints Photobooks apps and the industry leading SimplytoImpress card and stationery site, as well as Personal Creations, CafePress and ISeeMe! Visit to learn more about our brands. We have more than 500 team members across multiple offices, primarily in Calabasas CA, San Diego CA, Woodridge IL, Minneapolis, MN and Pleasanton, CA. We also have team members in two company-owned offices in China, as well as in Europe. Job Overview PlanetArt is seeking a Principal Staff Engineer-Web Platform to serve as a senior technical leader within our engineering organization and a key partner to the VP of Engineering. This is a highly hands-on role focused on building, operating, and scaling high-traffic ecommerce web platforms in a fast-moving production environment. Reporting directly to the VP of Engineering, this engineer will play a critical role in architecting, developing, troubleshooting, and maintaining our LAMP-based web applications and AWS infrastructure. The ideal candidate is equally comfortable writing production code, diagnosing complex site reliability issues, managing cloud infrastructure, and leading technical problem-solving during high-severity incidents. This role requires strong operational judgment and the ability to independently own production challenges across application, database, infrastructure, and deployment layers. The engineer will collaborate closely with our China-based development organization, serving as a senior US-based technical lead responsible for cross-team coordination, code quality, architectural guidance, and production stability. This is an ideal opportunity for an experienced engineer who thrives in high-scale ecommerce environments and enjoys combining deep application engineering with modern cloud operations and production ownership. PLEASE NOTE: Candidates much be local to or willing to relocate to the Calabasas area as we operate on a hybrid work model (3 days onsite, 2 remote) What You'll Do Key Responsibilities Full-Stack Platform Engineering: Design, develop, and maintain scalable features and services across our LAMP-based ecommerce platform, with a strong focus on reliability, performance, maintainability, and operational excellence. Production Operations & Incident Response : Act as a senior technical escalation point for complex production incidents, troubleshooting issues across application, infrastructure, networking, database, CDN, and deployment layers. Lead root cause analysis and drive long-term stability improvements. AWS Infrastructure Ownership : Manage and optimize AWS infrastructure, including deployment architecture, scaling strategies, observability, security, disaster recovery, and cost efficiency. Partner closely with DevOps and engineering leadership on operational best practices. High-Scale Performance Optimization : Monitor and improve application, database, and infrastructure performance for high-traffic consumer web applications. Identify bottlenecks and implement scalable solutions to improve uptime, latency, and system resilience. Cross-Functional Technical Leadership: Partner with Product, Design, Operations, and Customer Experience teams to translate business requirements into scalable technical solutions and ensure successful project execution. Global Engineering Collaboration : Work closely with the China-based engineering team to coordinate development efforts, conduct code reviews, align on architectural direction, manage releases, and maintain strong engineering communication across time zones. Code Quality & Engineering Standards : Champion high engineering standards through code reviews, testing strategies, documentation, observability, and operational best practices. Drive continuous improvement in system reliability and development processes. Technical Mentorship & Leadership : Provide technical mentorship and architectural guidance across the engineering organization. Influence technical direction through hands-on leadership, strong execution, and collaborative problem-solving. Requirements What You Should Have Skills, Qualifications, and Requirements Senior-Level Full-Stack Engineering Experience: 5+ years of professional experience building and operating large-scale web applications, including substantial hands-on experience with the LAMP stack (Linux, Apache, MySQL, PHP). Strong AWS & Cloud Operations Expertise : Deep hands-on experience with AWS services and production cloud environments, including EC2, RDS, S3, Lambda, CloudWatch, networking, scaling, monitoring, and infrastructure troubleshooting. Ecommerce & High-Traffic Website Experience : Experience supporting high-volume consumer-facing websites or ecommerce platforms, with a strong understanding of scalability, uptime, performance optimization, and operational reliability. Production Troubleshooting Expertise : Demonstrated ability to diagnose and resolve complex production issues under pressure, including database replication issues, performance degradation, infrastructure failures, deployment issues, and site outages. Distributed Systems & Database Knowledge : Strong understanding of distributed web architectures, database performance tuning, replication strategies, caching, queuing systems, and fault-tolerant system design. Global Team Collaboration: Experience working effectively with offshore or globally distributed engineering teams, with strong communication, coordination, and cross-cultural collaboration skills. Chinese Language Skills: Ability to communicate in Mandarin (spoken or written) is highly desirable to facilitate collaboration with our China-based engineering team. Engineering Best Practices: Strong understanding of software engineering fundamentals including Git workflows, CI/CD pipelines, automated testing, observability, code review practices, and secure development standards. Ownership Mentality: Self-directed engineer with strong operational instincts, excellent judgment, and the ability to independently own critical technical initiatives from design through production support. Technical Leadership: Demonstrated ability to influence engineering direction, mentor developers, and drive technical excellence through hands-on leadership rather than direct people management. What You Can Expect Working Conditions Work is performed in an office environment with low to moderate noise levels. Position requires regular, continuous use of computer. Position requires regular sitting and standing. Position requires regular interaction with team members through the following methods: in-person, phone, Zoom, Slack, or email. May require occasional travel. This is a hybrid position; employees are expected to be in the office three days per week (Monday, Tuesday, and Thursday) with the option of working remotely two days (Wednesday and Friday). Benefits The compensation range for this position is $130,000-$220,000 annual salary. PlanetArt offers a comprehensive benefits package, including: Health, Dental, and Vision Insurance Life Insurance Pet Insurance Mental Health Insurance 401(k) with matching Comprehensive Time Off Program including Paid Time Off, Sick Days, Paid Holidays, and Floating Holidays Employee Product Discounts
Principal DevOps Engineer
Ursa Major Berthoud, Colorado
Job Description Job Description The future of aerospace and defense starts here. Ursa Major was founded to revolutionize how America and its allies access and apply high-performance propulsion, from hypersonics to solid rocket motors, satellite maneuvering and launch. We design and deliver propulsion and defense systems that solve the most urgent and critical national security demands. Ursa Major is bringing a new model to propulsion and defense technology: one where performance and access are no longer constrained by legacy systems. We design, build, and deploy advanced propulsion systems that power national security missions, hypersonic capabilities, and the future of space access. Our mission requires an extraordinary team-one that will mold tomorrow's critical technologies while deploying today's best. We are an intrinsically motivated team with a passion for solving complex technical problems and empowering each other every day, knowing that there is always room for growth. As a Principal DevOps Engineer within our Digital Operations team, you won't just be managing servers or responding to deployment tickets. You will be a principal technical leader and system architect responsible for owning and shaping Ursa Major's platform engineering, infrastructure automation, and software deployment strategies across cloud, on-premise, and manufacturing environments. Sitting at the vital intersection of Software Engineering, IT, Zero-Trust Networking, Digital Manufacturing, and Cybersecurity, you will build the foundational platforms that empower our engineering and manufacturing teams to design, test, and deliver mission-critical systems faster and more reliably. As Ursa Major expands, you will decompose complex technical challenges, author clear architecture designs, and drive pragmatic, high-impact solutions that connect our cloud environments directly to physical manufacturing and test-bench hardware. Responsibilities: Platform Architecture & Developer Experience Lead the research, design, and continuous improvement of platform infrastructure, software deployment, and Developer Experience (DX) across the enterprise. Decompose complex infrastructure problems into maintainable systems; author clear technical design documents and lead cross-functional teams through execution. Act as a strategic connector between Software Engineering, IT, and Networking to unify tooling, share infrastructure patterns, and eliminate operational silos. Infrastructure-as-Code (IaC), GitOps & High Availability (HA) Advance our declarative Infrastructure-as-Code (IaC) and GitOps practices (using Terraform, Helm, and ArgoCD/Flux) to deliver fully automated, drift-free application lifecycles across Azure GovCloud and AWS GovCloud. Extend containerization and orchestration (Kubernetes, Docker) beyond cloud environments into on-premise test benches, shop-floor PCs, and edge manufacturing hardware in alignment with our CMMC Level 2 hybrid strategy. Own Disaster Recovery (DR) and Business Continuity (BC) architecture-maintaining automated backup and failover capabilities that keep critical system recovery times measured in minutes, not days. DevSecOps & AI Pipeline Integration Partner with Cybersecurity and Zero-Trust teams to integrate automated vulnerability scanning, secrets management, and policy-as-code guardrails directly into CI/CD workflows without introducing developer friction. Architect and support automated pipelines for hosted and private AI models, enabling software and engineering teams to leverage high-throughput AI capabilities safely within our secure compliance boundary. Maintain continuous telemetry, metrics, and alerting frameworks (Prometheus, Grafana) to give engineers real-time visibility into application performance and pipeline health. Leadership, Mentorship & Impact Frame all technical initiatives in terms of maximum impact on overall business health, execution speed, and system reliability. Mentor and elevate engineers across the software and IT organizations, fostering a culture of analytical, value-driven problem-solving. Function as a top-level technical decision-maker, managing high-ambiguity challenges with high autonomy while maintaining psychological safety and clear standards of excellence Minimum Qualifications: 8+ years of professional DevOps, SRE, or Infrastructure/Platform Engineering experience, with a proven track record of principal-level technical leadership and system architecture. Hands-on mastery of Azure GovCloud and/or AWS GovCloud environments, with deep expertise in VPC networking, IAM, and cloud-native security controls. Advanced proficiency in Terraform, Docker, Kubernetes, Helm, and continuous delivery via GitOps patterns (ArgoCD, GitHub Actions, or Flux). Direct experience architecting high-availability systems, automated DR solutions, and telemetry suites (Prometheus, Grafana). Practical experience configuring automated container/code and centralized secrets management (AWS/Azure Key Vaults, HashiCorp Vault). Experience automating, deploying, or supporting hosted AI model integrations within production CI/CD pipelines. A proactive problem solver who builds trust across teams, rejects black-and-white thinking, and focuses on enabling business outcomes. Preferred Qualifications: Familiarity with NIST SP 800-171 / CMMC Level 2 compliance standards and securing CUI data flows. Experience with Artifactory, SonarQube, Open Policy Agent (OPA), Kyverno, or automated Software Bill of Materials (SBOM) tooling. Experience automating physical hardware, edge nodes, or shop-floor environments using Packer, Ansible, K3s, or Rancher. Experience orchestrating self-hosted or private AI model execution frameworks. Ability to obtain and maintain a U.S. Government Security Clearance. Colorado law requires us to tell you the base compensation range of this role, which is $175,000 - $215,000, determined by your education, experience, knowledge, skills, and abilities. The salary range for this role is intentionally wide as we are evaluating individuals based on their unique experience and abilities to fit our needs. Most importantly, we are excited to meet you and see if you are a great fit for our team. What we can't quantify for you are the exciting challenges, supportive team, and amazing culture we enjoy. Classification: Full-Time, Exempt Benefits Include: (Please note, Interns are not eligible for benefits) Unlimited PTO - Vacation, Sick, Personal, and Bereavement Paid Parental and Adoptive Leave Medical, Dental and Vision Insurance Tax Advantage Accounts (HSA/FSA) Employer Paid Short and Long Term Disability, Basic Life, AD&D Additional Benefit Options Including Voluntary Life and Emergency Medical Transport EAP Program Retirement Savings Plan - 401k with Company Match Equity Grants in the Company How To Apply: Interested candidates are encouraged to apply by filling out the application below and clicking "Submit Application". This position will be posted for a minimum of 3 days and will remain open until filled or adjusted based on the volume of applicants. NOTE: Research suggests that women and BIPOC individuals may self-select out of opportunities if they don't meet 100% of the job requirements. We encourage anyone who believes they have the skills and the drive necessary to succeed here to apply for this role. Must be a U.S. Person (this includes U.S. Citizens and Permanent Residents). Eligibility to obtain and maintain a U.S. Security Clearance. We're an equal-opportunity employer. You will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, or disability status. No outside recruiters, please.
09/15/2026
Full time
Job Description Job Description The future of aerospace and defense starts here. Ursa Major was founded to revolutionize how America and its allies access and apply high-performance propulsion, from hypersonics to solid rocket motors, satellite maneuvering and launch. We design and deliver propulsion and defense systems that solve the most urgent and critical national security demands. Ursa Major is bringing a new model to propulsion and defense technology: one where performance and access are no longer constrained by legacy systems. We design, build, and deploy advanced propulsion systems that power national security missions, hypersonic capabilities, and the future of space access. Our mission requires an extraordinary team-one that will mold tomorrow's critical technologies while deploying today's best. We are an intrinsically motivated team with a passion for solving complex technical problems and empowering each other every day, knowing that there is always room for growth. As a Principal DevOps Engineer within our Digital Operations team, you won't just be managing servers or responding to deployment tickets. You will be a principal technical leader and system architect responsible for owning and shaping Ursa Major's platform engineering, infrastructure automation, and software deployment strategies across cloud, on-premise, and manufacturing environments. Sitting at the vital intersection of Software Engineering, IT, Zero-Trust Networking, Digital Manufacturing, and Cybersecurity, you will build the foundational platforms that empower our engineering and manufacturing teams to design, test, and deliver mission-critical systems faster and more reliably. As Ursa Major expands, you will decompose complex technical challenges, author clear architecture designs, and drive pragmatic, high-impact solutions that connect our cloud environments directly to physical manufacturing and test-bench hardware. Responsibilities: Platform Architecture & Developer Experience Lead the research, design, and continuous improvement of platform infrastructure, software deployment, and Developer Experience (DX) across the enterprise. Decompose complex infrastructure problems into maintainable systems; author clear technical design documents and lead cross-functional teams through execution. Act as a strategic connector between Software Engineering, IT, and Networking to unify tooling, share infrastructure patterns, and eliminate operational silos. Infrastructure-as-Code (IaC), GitOps & High Availability (HA) Advance our declarative Infrastructure-as-Code (IaC) and GitOps practices (using Terraform, Helm, and ArgoCD/Flux) to deliver fully automated, drift-free application lifecycles across Azure GovCloud and AWS GovCloud. Extend containerization and orchestration (Kubernetes, Docker) beyond cloud environments into on-premise test benches, shop-floor PCs, and edge manufacturing hardware in alignment with our CMMC Level 2 hybrid strategy. Own Disaster Recovery (DR) and Business Continuity (BC) architecture-maintaining automated backup and failover capabilities that keep critical system recovery times measured in minutes, not days. DevSecOps & AI Pipeline Integration Partner with Cybersecurity and Zero-Trust teams to integrate automated vulnerability scanning, secrets management, and policy-as-code guardrails directly into CI/CD workflows without introducing developer friction. Architect and support automated pipelines for hosted and private AI models, enabling software and engineering teams to leverage high-throughput AI capabilities safely within our secure compliance boundary. Maintain continuous telemetry, metrics, and alerting frameworks (Prometheus, Grafana) to give engineers real-time visibility into application performance and pipeline health. Leadership, Mentorship & Impact Frame all technical initiatives in terms of maximum impact on overall business health, execution speed, and system reliability. Mentor and elevate engineers across the software and IT organizations, fostering a culture of analytical, value-driven problem-solving. Function as a top-level technical decision-maker, managing high-ambiguity challenges with high autonomy while maintaining psychological safety and clear standards of excellence Minimum Qualifications: 8+ years of professional DevOps, SRE, or Infrastructure/Platform Engineering experience, with a proven track record of principal-level technical leadership and system architecture. Hands-on mastery of Azure GovCloud and/or AWS GovCloud environments, with deep expertise in VPC networking, IAM, and cloud-native security controls. Advanced proficiency in Terraform, Docker, Kubernetes, Helm, and continuous delivery via GitOps patterns (ArgoCD, GitHub Actions, or Flux). Direct experience architecting high-availability systems, automated DR solutions, and telemetry suites (Prometheus, Grafana). Practical experience configuring automated container/code and centralized secrets management (AWS/Azure Key Vaults, HashiCorp Vault). Experience automating, deploying, or supporting hosted AI model integrations within production CI/CD pipelines. A proactive problem solver who builds trust across teams, rejects black-and-white thinking, and focuses on enabling business outcomes. Preferred Qualifications: Familiarity with NIST SP 800-171 / CMMC Level 2 compliance standards and securing CUI data flows. Experience with Artifactory, SonarQube, Open Policy Agent (OPA), Kyverno, or automated Software Bill of Materials (SBOM) tooling. Experience automating physical hardware, edge nodes, or shop-floor environments using Packer, Ansible, K3s, or Rancher. Experience orchestrating self-hosted or private AI model execution frameworks. Ability to obtain and maintain a U.S. Government Security Clearance. Colorado law requires us to tell you the base compensation range of this role, which is $175,000 - $215,000, determined by your education, experience, knowledge, skills, and abilities. The salary range for this role is intentionally wide as we are evaluating individuals based on their unique experience and abilities to fit our needs. Most importantly, we are excited to meet you and see if you are a great fit for our team. What we can't quantify for you are the exciting challenges, supportive team, and amazing culture we enjoy. Classification: Full-Time, Exempt Benefits Include: (Please note, Interns are not eligible for benefits) Unlimited PTO - Vacation, Sick, Personal, and Bereavement Paid Parental and Adoptive Leave Medical, Dental and Vision Insurance Tax Advantage Accounts (HSA/FSA) Employer Paid Short and Long Term Disability, Basic Life, AD&D Additional Benefit Options Including Voluntary Life and Emergency Medical Transport EAP Program Retirement Savings Plan - 401k with Company Match Equity Grants in the Company How To Apply: Interested candidates are encouraged to apply by filling out the application below and clicking "Submit Application". This position will be posted for a minimum of 3 days and will remain open until filled or adjusted based on the volume of applicants. NOTE: Research suggests that women and BIPOC individuals may self-select out of opportunities if they don't meet 100% of the job requirements. We encourage anyone who believes they have the skills and the drive necessary to succeed here to apply for this role. Must be a U.S. Person (this includes U.S. Citizens and Permanent Residents). Eligibility to obtain and maintain a U.S. Security Clearance. We're an equal-opportunity employer. You will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, or disability status. No outside recruiters, please.
Staff / Principal Platform Engineer
AppGate Cybersecurity, Inc. New York, New York
Job Description Job Description Staff / Principal Platform Engineer Location: New York City Hybrid Department: AI Platform & Infrastructure Team Reports to: Vangie Shue - Principal Engineering Manager About AppGate AppGate secures and protects an organization's most valuable assets with its high performance Zero Trust Network Access (ZTNA) solution and Cyber Advisory Services. AppGate ZTNA is the only direct-routed Zero Trust solution built for peak performance, superior protection and seamless interoperability. AppGate Cyber Advisory Services harden your security posture and ensure business continuity. AppGate safeguards Fortune 500 enterprises and government agencies worldwide. Learn more at About the Role As we expand our platform, we are standing up a new AI Platform & Infrastructure team: the engine room of AppGate's AI strategy. This team owns the infrastructure layer that every next-generation security capability is built on, from network observability to AI-driven threat detection and the secure operation of emerging Agentic AI systems. We're looking for a Staff or Principal Platform Engineer to build and operate the foundational platform behind AppGate's AI products. You combine deep DevOps and cloud infrastructure expertise with hands-on experience operationalizing AI/ML systems, and you treat observability as a first-class engineering discipline. This is a rare opportunity to join a small, private, high-impact company where your work directly shapes the architecture, reliability and core platform that defines the future of security. You'll own the platform spanning APIs, cloud and self-managed solutions and AI/ML infrastructure, and you'll make it fast, reliable and observable at scale. This is a high-leverage, hands-on role for a senior engineer who sets technical direction and still ships. Key Responsibilities Build the Platform: design, build and operate the cloud infrastructure, services and pipelines that AppGate's AI and cloud products run on. Strong experience with self-managed technologies (kafka, elasticsearch) and Kubernetes are a must. Infrastructure as Code & Deployment Orchestration: Terraform and Helm for cloud provisioning, service deployment and configuration management. Implement Observability: instrument APIs, cloud services and AI/ML infrastructure with metrics, logging, tracing and alerting, and define SLOs and operational health metrics that teams trust. Data Platform: real-time and batch data ingestion pipelines, feature stores and data quality. Integrations: third-party connectors, APIs and platform integrations. Operationalize AI/ML: build model serving and inference pipelines, experiment tracking and the MLOps tooling for deployment, versioning, drift monitoring and lifecycle management. Engineer for reliability & automation: apply SRE practices to reduce toil, improve resilience and keep latency and uptime within target across the platform. Automate everything - deliver infrastructure-as-code, CI/CD and self-service tooling so product teams ship safely and quickly. Set technical direction: define platform standards, architecture and best practices, and raise the engineering bar through design reviews and mentorship. Collaborate cross-functionally: partner with data scientists, product teams and leadership to align platform investment with AppGate's strategic vision. Required Qualifications Experience: extensive platform, infrastructure or SRE engineering experience, with a track record of operating production systems at scale. Staff-level candidates typically bring 8+ years and Principal-level candidates 12+ years, though we hire on demonstrated impact. DevOps depth: strong command of infrastructure-as-code (Terraform or equivalent), CI/CD, containers and orchestration (Docker, Kubernetes), and cloud platforms (AWS). Observability expertise: hands-on experience implementing observability across APIs, cloud services and distributed systems using tools such as Prometheus, Grafana, OpenTelemetry, the ELK stack or comparable, including SLO and error-budget practice. Data platform skills: familiarity with real-time and batch ingestion pipelines, feature stores and data quality at production scale. Engineering craft: fluency in a primary backend language (Python, Go or similar) and a strong bias toward automation, testing and reliable, maintainable systems. Leadership: a record of setting technical direction, leading complex initiatives across teams, mentoring senior engineers, while still being very hands-on. Mindset: pragmatic, rigorous and ownership-driven. You thrive in a small, fast-moving environment and enjoy building foundations others depend on. Preferred Qualifications AI/ML infrastructure: experience building or operating model serving, inference pipelines and MLOps tooling such as MLflow, Kubeflow, SageMaker or equivalent, including model deployment, versioning and drift monitoring. Networking & Zero Trust fundamentals: working knowledge of the network and routing layer beneath modern access solutions - TCP/IP, TLS, tunneling/overlay networks, packet routing and filtering, DNS and firewalling - and familiarity with Zero Trust Network Access (ZTNA) or adjacent domains (VPN, SDP, SASE, software-defined networking). You can reason about traffic paths, latency and throughput end-to-end, and instrument the network as a first-class observability signal. Compensation Staff: 185k-225k base Principal: 215k-270k base We offer performance bonuses and considerable equity. AppGate is An Equal Opportunity/Affirmative Action Employer and a federal contractor subject to the Rehabilitation Act of 1973 and the Vietnam Era Veterans Readjustment Assistance Act of 1974 as amended, and their corresponding regulations. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or veteran status, age or any other federally protected class. Further, AppGate is an affirmative action employer committed to taking positive steps to employ, advance in employment and otherwise afford equal employment opportunity to protected veterans and individuals with disabilities. In furtherance of AppGate's policy regarding affirmative action and equal employment opportunity, AppGate has developed a written affirmative action program. This program is available for review upon request by any applicant or employee during normal business hours by contacting the company's EEO Coordinator.
09/15/2026
Full time
Job Description Job Description Staff / Principal Platform Engineer Location: New York City Hybrid Department: AI Platform & Infrastructure Team Reports to: Vangie Shue - Principal Engineering Manager About AppGate AppGate secures and protects an organization's most valuable assets with its high performance Zero Trust Network Access (ZTNA) solution and Cyber Advisory Services. AppGate ZTNA is the only direct-routed Zero Trust solution built for peak performance, superior protection and seamless interoperability. AppGate Cyber Advisory Services harden your security posture and ensure business continuity. AppGate safeguards Fortune 500 enterprises and government agencies worldwide. Learn more at About the Role As we expand our platform, we are standing up a new AI Platform & Infrastructure team: the engine room of AppGate's AI strategy. This team owns the infrastructure layer that every next-generation security capability is built on, from network observability to AI-driven threat detection and the secure operation of emerging Agentic AI systems. We're looking for a Staff or Principal Platform Engineer to build and operate the foundational platform behind AppGate's AI products. You combine deep DevOps and cloud infrastructure expertise with hands-on experience operationalizing AI/ML systems, and you treat observability as a first-class engineering discipline. This is a rare opportunity to join a small, private, high-impact company where your work directly shapes the architecture, reliability and core platform that defines the future of security. You'll own the platform spanning APIs, cloud and self-managed solutions and AI/ML infrastructure, and you'll make it fast, reliable and observable at scale. This is a high-leverage, hands-on role for a senior engineer who sets technical direction and still ships. Key Responsibilities Build the Platform: design, build and operate the cloud infrastructure, services and pipelines that AppGate's AI and cloud products run on. Strong experience with self-managed technologies (kafka, elasticsearch) and Kubernetes are a must. Infrastructure as Code & Deployment Orchestration: Terraform and Helm for cloud provisioning, service deployment and configuration management. Implement Observability: instrument APIs, cloud services and AI/ML infrastructure with metrics, logging, tracing and alerting, and define SLOs and operational health metrics that teams trust. Data Platform: real-time and batch data ingestion pipelines, feature stores and data quality. Integrations: third-party connectors, APIs and platform integrations. Operationalize AI/ML: build model serving and inference pipelines, experiment tracking and the MLOps tooling for deployment, versioning, drift monitoring and lifecycle management. Engineer for reliability & automation: apply SRE practices to reduce toil, improve resilience and keep latency and uptime within target across the platform. Automate everything - deliver infrastructure-as-code, CI/CD and self-service tooling so product teams ship safely and quickly. Set technical direction: define platform standards, architecture and best practices, and raise the engineering bar through design reviews and mentorship. Collaborate cross-functionally: partner with data scientists, product teams and leadership to align platform investment with AppGate's strategic vision. Required Qualifications Experience: extensive platform, infrastructure or SRE engineering experience, with a track record of operating production systems at scale. Staff-level candidates typically bring 8+ years and Principal-level candidates 12+ years, though we hire on demonstrated impact. DevOps depth: strong command of infrastructure-as-code (Terraform or equivalent), CI/CD, containers and orchestration (Docker, Kubernetes), and cloud platforms (AWS). Observability expertise: hands-on experience implementing observability across APIs, cloud services and distributed systems using tools such as Prometheus, Grafana, OpenTelemetry, the ELK stack or comparable, including SLO and error-budget practice. Data platform skills: familiarity with real-time and batch ingestion pipelines, feature stores and data quality at production scale. Engineering craft: fluency in a primary backend language (Python, Go or similar) and a strong bias toward automation, testing and reliable, maintainable systems. Leadership: a record of setting technical direction, leading complex initiatives across teams, mentoring senior engineers, while still being very hands-on. Mindset: pragmatic, rigorous and ownership-driven. You thrive in a small, fast-moving environment and enjoy building foundations others depend on. Preferred Qualifications AI/ML infrastructure: experience building or operating model serving, inference pipelines and MLOps tooling such as MLflow, Kubeflow, SageMaker or equivalent, including model deployment, versioning and drift monitoring. Networking & Zero Trust fundamentals: working knowledge of the network and routing layer beneath modern access solutions - TCP/IP, TLS, tunneling/overlay networks, packet routing and filtering, DNS and firewalling - and familiarity with Zero Trust Network Access (ZTNA) or adjacent domains (VPN, SDP, SASE, software-defined networking). You can reason about traffic paths, latency and throughput end-to-end, and instrument the network as a first-class observability signal. Compensation Staff: 185k-225k base Principal: 215k-270k base We offer performance bonuses and considerable equity. AppGate is An Equal Opportunity/Affirmative Action Employer and a federal contractor subject to the Rehabilitation Act of 1973 and the Vietnam Era Veterans Readjustment Assistance Act of 1974 as amended, and their corresponding regulations. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or veteran status, age or any other federally protected class. Further, AppGate is an affirmative action employer committed to taking positive steps to employ, advance in employment and otherwise afford equal employment opportunity to protected veterans and individuals with disabilities. In furtherance of AppGate's policy regarding affirmative action and equal employment opportunity, AppGate has developed a written affirmative action program. This program is available for review upon request by any applicant or employee during normal business hours by contacting the company's EEO Coordinator.
Principal Software Engineer, Platform Engineering & Edge Infrastructure
Saviynt Milpitas, California
Job Description Job Description Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world's leading brands, Fortune 500 companies and government institutions. For more information, please visit . Come join us as founding members of Saviynt's AI Security team and help us build out AI security for the world's leading enterprises. WHAT YOU WILL BE DOING Design and operate the infrastructure powering Cloud and Edge platform deployments. Build deployment automation for distributed Edge PoPs supporting headquarters, branch offices, regional hubs, and customer data centers. Design highly available deployment architectures supporting secure fail-closed operation and resilient policy synchronization. Build CI/CD pipelines, release automation, upgrade orchestration, and lifecycle management for Cloud, Edge, and Endpoint components. Develop observability platforms, including logging, metrics, tracing, health monitoring, and audit pipelines. Automate provisioning, certificate lifecycle management, secrets management, and secure configuration distribution. Drive platform scalability, operational excellence, reliability, disaster recovery, and security. AI & Agentic Engineering Apply AI-assisted engineering across infrastructure, deployment automation, and platform operations. Build infrastructure supporting AI-native applications, AI services, and distributed agentic workloads. Follow AI SDLC best practices for software delivery, automation, deployment, monitoring, and continuous improvement. Evaluate emerging AI infrastructure technologies to improve engineering productivity and operational efficiency. WHAT YOU BRING 1+ years of Principal-level of platform engineering, DevOps, or Site Reliability Engineering experience. Deep experience operating Kubernetes and cloud-native platforms at enterprise scale. Strong experience with Terraform, Helm, GitHub Actions, ArgoCD, Ansible, or similar automation technologies. Experience with distributed networking, service meshes, proxies, DNS, load balancing, and TLS. Strong experience in Linux systems administration and infrastructure automation. Experience deploying highly available distributed enterprise software across multiple customer environments. Hands-on experience using AI-assisted development tools such as GitHub Copilot, Cursor, Claude Code, Windsurf, ChatGPT, or similar. Understanding of AI application architectures, AI agents, MCP, and AI-enabled infrastructure. Familiarity with AI SDLC best practices, including AI-assisted development, automated testing, CI/CD automation, observability, and responsible use of AI-generated code. Strong operational mindset with excellent communication, collaboration, and leadership skills. We offer you a competitive total rewards package, learning and tremendous opportunities to grow and advance in your career. At Saviynt, it is not typical for an individual to be hired at or near the top of the range for their role and final compensation decisions are dependent on many factors including but are not limited to location; skill sets; experience and training; licensure and certifications; and other relevant business and organizational needs. A reasonable estimate of the current range is $240,000 - $250,000 annually. You may also be eligible to participate in a Saviynt discretionary bonus plan, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance. Saviynt is an amazing place to work. We are a high-growth, Platform as a Service company focused on Identity Authority to power and protect the world at work. You will experience tremendous growth and learning opportunities through challenging yet rewarding work which directly impacts our customers, all within a welcoming and positive work environment. If you're resilient and enjoy working in a dynamic environment you belong with us! Security & Compliance This role requires adherence to Saviynt's information security and privacy policies and procedures, including annual security training. Saviynt is an equal opportunity employer and we welcome everyone to our team. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
09/15/2026
Full time
Job Description Job Description Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world's leading brands, Fortune 500 companies and government institutions. For more information, please visit . Come join us as founding members of Saviynt's AI Security team and help us build out AI security for the world's leading enterprises. WHAT YOU WILL BE DOING Design and operate the infrastructure powering Cloud and Edge platform deployments. Build deployment automation for distributed Edge PoPs supporting headquarters, branch offices, regional hubs, and customer data centers. Design highly available deployment architectures supporting secure fail-closed operation and resilient policy synchronization. Build CI/CD pipelines, release automation, upgrade orchestration, and lifecycle management for Cloud, Edge, and Endpoint components. Develop observability platforms, including logging, metrics, tracing, health monitoring, and audit pipelines. Automate provisioning, certificate lifecycle management, secrets management, and secure configuration distribution. Drive platform scalability, operational excellence, reliability, disaster recovery, and security. AI & Agentic Engineering Apply AI-assisted engineering across infrastructure, deployment automation, and platform operations. Build infrastructure supporting AI-native applications, AI services, and distributed agentic workloads. Follow AI SDLC best practices for software delivery, automation, deployment, monitoring, and continuous improvement. Evaluate emerging AI infrastructure technologies to improve engineering productivity and operational efficiency. WHAT YOU BRING 1+ years of Principal-level of platform engineering, DevOps, or Site Reliability Engineering experience. Deep experience operating Kubernetes and cloud-native platforms at enterprise scale. Strong experience with Terraform, Helm, GitHub Actions, ArgoCD, Ansible, or similar automation technologies. Experience with distributed networking, service meshes, proxies, DNS, load balancing, and TLS. Strong experience in Linux systems administration and infrastructure automation. Experience deploying highly available distributed enterprise software across multiple customer environments. Hands-on experience using AI-assisted development tools such as GitHub Copilot, Cursor, Claude Code, Windsurf, ChatGPT, or similar. Understanding of AI application architectures, AI agents, MCP, and AI-enabled infrastructure. Familiarity with AI SDLC best practices, including AI-assisted development, automated testing, CI/CD automation, observability, and responsible use of AI-generated code. Strong operational mindset with excellent communication, collaboration, and leadership skills. We offer you a competitive total rewards package, learning and tremendous opportunities to grow and advance in your career. At Saviynt, it is not typical for an individual to be hired at or near the top of the range for their role and final compensation decisions are dependent on many factors including but are not limited to location; skill sets; experience and training; licensure and certifications; and other relevant business and organizational needs. A reasonable estimate of the current range is $240,000 - $250,000 annually. You may also be eligible to participate in a Saviynt discretionary bonus plan, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance. Saviynt is an amazing place to work. We are a high-growth, Platform as a Service company focused on Identity Authority to power and protect the world at work. You will experience tremendous growth and learning opportunities through challenging yet rewarding work which directly impacts our customers, all within a welcoming and positive work environment. If you're resilient and enjoy working in a dynamic environment you belong with us! Security & Compliance This role requires adherence to Saviynt's information security and privacy policies and procedures, including annual security training. Saviynt is an equal opportunity employer and we welcome everyone to our team. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Mastercard
Site Reliability Engineering Intern - Payments Technology
Mastercard O Fallon, Missouri
As a Site Reliability Engineering Intern at Mastercard, you will support the design, monitoring, and operation of large-scale, cloud-based systems that power secure, global payments. Working alongside experienced SREs and software engineers, you'll help maintain production environments, assist in incident response, and contribute to root-cause analysis to prevent recurrence. You'll aid in building automation for deployments, monitoring, and alerting, and help document runbooks and processes. This internship offers exposure to cutting-edge technologies, modern DevOps practices, and a collaborative culture focused on innovation and continuous learning. Responsibilities Assist in maintaining and monitoring production and cloud infrastructure Support incident response, troubleshooting, and root-cause analysis Help improve system reliability, scalability, and performance Contribute to automation of deployments, monitoring, and alerts Work with engineers to implement reliability best practices Document processes, runbooks, and technical findings Participate in on-call simulations and reliability exercises Collaborate with cross-functional teams in an Agile environment Required Skills Linux system administration basics Cloud computing fundamentals (AWS/GCP/Azure) Scripting/programming (Python, Java, or Go) CI/CD pipelines and Dev Ops concepts Containers and orchestration (Docker, Kubernetes) Monitoring and logging tools (Prometheus, Grafana, ELK, etc.) Networking fundamentals (TCP/IP, DNS, HTTP) Version control with Git Automation and Infrastructure as Code concepts Troubleshooting and root-cause analysis
09/13/2026
Full time
As a Site Reliability Engineering Intern at Mastercard, you will support the design, monitoring, and operation of large-scale, cloud-based systems that power secure, global payments. Working alongside experienced SREs and software engineers, you'll help maintain production environments, assist in incident response, and contribute to root-cause analysis to prevent recurrence. You'll aid in building automation for deployments, monitoring, and alerting, and help document runbooks and processes. This internship offers exposure to cutting-edge technologies, modern DevOps practices, and a collaborative culture focused on innovation and continuous learning. Responsibilities Assist in maintaining and monitoring production and cloud infrastructure Support incident response, troubleshooting, and root-cause analysis Help improve system reliability, scalability, and performance Contribute to automation of deployments, monitoring, and alerts Work with engineers to implement reliability best practices Document processes, runbooks, and technical findings Participate in on-call simulations and reliability exercises Collaborate with cross-functional teams in an Agile environment Required Skills Linux system administration basics Cloud computing fundamentals (AWS/GCP/Azure) Scripting/programming (Python, Java, or Go) CI/CD pipelines and Dev Ops concepts Containers and orchestration (Docker, Kubernetes) Monitoring and logging tools (Prometheus, Grafana, ELK, etc.) Networking fundamentals (TCP/IP, DNS, HTTP) Version control with Git Automation and Infrastructure as Code concepts Troubleshooting and root-cause analysis

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board