Job Description Job Description Overview Principal Data Links and Communications Engineer LOCATION : Hanscom AFB, MA Salary Range : $145-$155,000 annually depending on experience, certifications, and qualifications JOB STATUS: Full-time CLEARANCE : Top Secret/SCI TRAVEL: Limited, as needed Astrion has an exciting opportunity for a Principal Data Links and Communications Engineer for the EPASS Cyber and Networks Directorate (EPASS HN II) contract , supporting the Aerial Networks Division (AFLCMC/HNA) located at Hanscom AFB in Massachusetts . REQUIRED QUALIFICATIONS / SKILLS: Must be a US Citizen Must have and be able to maintain a Top Secret/SCI clearance Knowledge and experience with Link 22 and tatical data links engineering DESIRED QUALIFICATIONS / SKILLS: 5 Years experience in software design and software requirements development Proficiency in C++ an Java Programming languages and development frameworks Proficiency in MBSE modeling, languages, and tools (DODAF, SysML, UML, Cameo) Networking and embedded mission systems experience Familiarity with government acquisition planning, tailoring and documentation (Acquisition Strategy Plan, Statement of Work, System Engineering Plan, User Agreements, Product Roadmap). ADDITIONAL QUALIFICATIONS: Provide knowledge and experience in the development, integration, testing, and deployment of Air Force systems and their associated data links, to include definition of data link interoperability, interfaces, and verification and test assessment requirements. This includes interfacing with the Federal, Department of Defense (DoD), United States Air Force (USAF), United States Army, and Joint communities to ensure the data link communications capability complies with the applicable Federal Aviation Administration, Defense Information Systems Agency (DISA), National Security Agency (NSA), Service Acquisition Executive, Joint Mission Planning System (JMPS), Damage Assessment and Casualty Report business and technical rules. Provide resources with knowledge and experience in Developmental and Operational test and Evaluation at the sub-system and system level. This includes interface with Air Combat Command, Air Force Command and Control Integration Center, Joint Interoperability Test Command, Space and Naval Warfare Systems Command, and all levels of the organizations that develop and manage the C3 and encryption key distribution infrastructure for controlling a networked system, providing technical advice on C4 issues, including pre-launch conditioning and planning, aircraft-weapon pre- and post-release communications, and real-time targeting. The Contractor shall advise and assist the program Chief Engineer, Integrated Product Team (IPT) Leads, Program Managers (PMs), and joint combined test team customers on performance, schedule, and cost issues with respect to planning, communications, and control. The contractor will perform other duties as assigned. The candidate may be called upon for knowledge and experience in the following specific areas: Combat Net Radio (CNR) Test and Demonstration include engineering support for interoperability testing weapon configurations with Threshold and Objective launch platforms/controllers and ground Joint Terminal Attack Controllers over Ultra-High Frequency (UHF) CNR and Link 22 communications. Joint Concept of Employment (CONEMP) Cryptologic Systems Link 16 Networking Interface Control Digitally Aided Close Air Support (DACAS) United States Message Text Format (USMTF) Joint Interoperability of Tactical Command and Control Systems (JINTACCS) RESPONSIBILITIES: Support defining and documenting specifications for network-enabled weapon systems to ensure Link 16 J11.X message interoperability Interface with DoD and non-DoD federal agencies to ensure the data link communications capability complies with the applicable regulations Interface with Air Force, other services and agencies in support of developing communications equipment and managing encryption infrastructure for controlling a networked system Provide technical advice on Command, Control and Communications (C3) issues, including pre-launch conditioning and planning, networked weapons, and real-time targeting Advise and assist the staff Engineers, IPT members, PMs, and test team customers on performance, schedule, and cost issues with respect to planning, communications, and control Support defining and documenting a set of Combat Net Radio (CNR) networking parameter configurations to provide the best balance between flexible implementation of capability and implementation impact on both existing and emerging weapon systems Provide engineering support for interoperability testing weapon configurations over UHF CNR and Link 16 Support developing, documenting, and presenting a Joint Concept of Employment (CONEMP) which is flexible, scalable, and compatible with communications via Military tactical data exchange network Support developing and documenting Joint Service approaches for integrating weapons into existing TDL environments Support development and documentation of Link 16 J11.X messages within the construct of a Joint Service CM strategy that will maintain interoperability between weapon systems and controllers while supporting a flexible CONEMP
09/20/2026
Full time
Job Description Job Description Overview Principal Data Links and Communications Engineer LOCATION : Hanscom AFB, MA Salary Range : $145-$155,000 annually depending on experience, certifications, and qualifications JOB STATUS: Full-time CLEARANCE : Top Secret/SCI TRAVEL: Limited, as needed Astrion has an exciting opportunity for a Principal Data Links and Communications Engineer for the EPASS Cyber and Networks Directorate (EPASS HN II) contract , supporting the Aerial Networks Division (AFLCMC/HNA) located at Hanscom AFB in Massachusetts . REQUIRED QUALIFICATIONS / SKILLS: Must be a US Citizen Must have and be able to maintain a Top Secret/SCI clearance Knowledge and experience with Link 22 and tatical data links engineering DESIRED QUALIFICATIONS / SKILLS: 5 Years experience in software design and software requirements development Proficiency in C++ an Java Programming languages and development frameworks Proficiency in MBSE modeling, languages, and tools (DODAF, SysML, UML, Cameo) Networking and embedded mission systems experience Familiarity with government acquisition planning, tailoring and documentation (Acquisition Strategy Plan, Statement of Work, System Engineering Plan, User Agreements, Product Roadmap). ADDITIONAL QUALIFICATIONS: Provide knowledge and experience in the development, integration, testing, and deployment of Air Force systems and their associated data links, to include definition of data link interoperability, interfaces, and verification and test assessment requirements. This includes interfacing with the Federal, Department of Defense (DoD), United States Air Force (USAF), United States Army, and Joint communities to ensure the data link communications capability complies with the applicable Federal Aviation Administration, Defense Information Systems Agency (DISA), National Security Agency (NSA), Service Acquisition Executive, Joint Mission Planning System (JMPS), Damage Assessment and Casualty Report business and technical rules. Provide resources with knowledge and experience in Developmental and Operational test and Evaluation at the sub-system and system level. This includes interface with Air Combat Command, Air Force Command and Control Integration Center, Joint Interoperability Test Command, Space and Naval Warfare Systems Command, and all levels of the organizations that develop and manage the C3 and encryption key distribution infrastructure for controlling a networked system, providing technical advice on C4 issues, including pre-launch conditioning and planning, aircraft-weapon pre- and post-release communications, and real-time targeting. The Contractor shall advise and assist the program Chief Engineer, Integrated Product Team (IPT) Leads, Program Managers (PMs), and joint combined test team customers on performance, schedule, and cost issues with respect to planning, communications, and control. The contractor will perform other duties as assigned. The candidate may be called upon for knowledge and experience in the following specific areas: Combat Net Radio (CNR) Test and Demonstration include engineering support for interoperability testing weapon configurations with Threshold and Objective launch platforms/controllers and ground Joint Terminal Attack Controllers over Ultra-High Frequency (UHF) CNR and Link 22 communications. Joint Concept of Employment (CONEMP) Cryptologic Systems Link 16 Networking Interface Control Digitally Aided Close Air Support (DACAS) United States Message Text Format (USMTF) Joint Interoperability of Tactical Command and Control Systems (JINTACCS) RESPONSIBILITIES: Support defining and documenting specifications for network-enabled weapon systems to ensure Link 16 J11.X message interoperability Interface with DoD and non-DoD federal agencies to ensure the data link communications capability complies with the applicable regulations Interface with Air Force, other services and agencies in support of developing communications equipment and managing encryption infrastructure for controlling a networked system Provide technical advice on Command, Control and Communications (C3) issues, including pre-launch conditioning and planning, networked weapons, and real-time targeting Advise and assist the staff Engineers, IPT members, PMs, and test team customers on performance, schedule, and cost issues with respect to planning, communications, and control Support defining and documenting a set of Combat Net Radio (CNR) networking parameter configurations to provide the best balance between flexible implementation of capability and implementation impact on both existing and emerging weapon systems Provide engineering support for interoperability testing weapon configurations over UHF CNR and Link 16 Support developing, documenting, and presenting a Joint Concept of Employment (CONEMP) which is flexible, scalable, and compatible with communications via Military tactical data exchange network Support developing and documenting Joint Service approaches for integrating weapons into existing TDL environments Support development and documentation of Link 16 J11.X messages within the construct of a Joint Service CM strategy that will maintain interoperability between weapon systems and controllers while supporting a flexible CONEMP
Job Description Job Description Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you'll be part of a passionate team dedicated to accomplishing hard things, together. Principal ML Platform Engineer Lead the engineering of software that matters - driving AI automation for the world's largest enterprises. This role is based at our h eadquarters in McLean, Virginia. Appian was built on a culture of in-person collaboration, which we believe is a key driver of our mission to be the best. Employees hired for this position are expected to be in the office 5 days a week to foster that culture and ensure we continue to thrive through shared ideas and teamwork. We believe being in the office provides more opportunities to come together and celebrate working with the exceptional people across Appian. About the Team Appian Engineering spans the full depth of our platform: from the foundational layers that power enterprise scale, to the AI capabilities redefining what automation can do. We operate in a highly collaborative, fast-paced environment focused on technical precision, continuous learning, and high code quality. By joining our team, you will solve real-world problems that directly shape how Appian delivers our AI-Powered Process Automation platform to enterprises around the world. The Opportunity As a Principal Software Engineer, you will serve as a technical linchpin for the team, bringing deep expertise in cloud-native architecture and a track record of influencing engineering direction beyond your immediate scope. Equipped with cutting-edge AI tooling, you will drive the design and delivery of high-complexity engineering solutions, set the technical bar for the team, and lead other engineers towards solutions of real complexity to ensure flawless Enterprise-Grade Orchestration. What You'll Do Develop Clean Software: Architect, build, and optimize high-performance software systems while maintaining a strong personal technical presence on the team. Lead Platform Modernization: Spearhead strategic technological changes and champion code refactoring efforts to keep the core Appian codebase cutting-edge, modern, and performant. Engineer with AI: Use AI coding tools fluently as a force multiplier: generating, reviewing, and critically evaluating AI-assisted code to ship faster without compromising quality or correctness. Lead Architecture & Delivery: Drive technical story breakdowns, acceptance criteria, and architectural design across complex, multi-tier application layers - from feature scoping through implementation. Optimize Performance & Scale: Manage product availability, latency, scalability, and efficiency by engineering deep reliability into our core software systems and performing advanced system tuning. Drive Engineering Excellence: Radiate development best practices across the department, perform meticulous code reviews on design and implementation, and build automation frameworks to prevent problem recurrence. Lead & Grow Engineers: Actively coach and mentor engineers at multiple levels, identify and close skill gaps on the team, and take ownership of accelerating the technical growth of those around you. Influence Technical Documentation: Share your expert domain knowledge regularly across the department, building a reputation as a vital resource and publishing high-quality content to Engineering's permanent documentation site. Required Qualifications Education: Minimum of a Bachelor of Science degree in Computer Science or a related technical/analytical discipline. (Equivalent experience is not accepted in lieu of a degree). Experience: 10+ years of relevant software development experience with a BS (or 8+ years of experience paired with a Master of Science in Computer Science or related field). Technical Mastery: Expert coding, scripting, and debugging proficiency in one or more core enterprise programming languages, specifically Java, Python, or Go. Domain Expertise: Deep working knowledge of distributed systems, cloud infrastructure, and the ability to contribute meaningfully at a senior individual contributor level within that space. Cross-Team Influence: Demonstrated ability to drive technical decisions and shape engineering practices beyond a single team or project scope. AI-Augmented Development: Demonstrated experience using AI coding assistants and a strong ability to evaluate, coach others on, and selectively apply AI-generated code in a production engineering context. Production Mastery: Proven experience developing, optimizing, and maintaining a high-volume, mission-critical production service environment. Communication & Alignment: Exceptional ability to communicate highly technical architectures verbally, visually, and in writing to diverse engineering audiences. Preferred Qualifications Cloud Architecture: Strong experience designing microservices, working with containerization (Docker, Kubernetes), and implementing modern CI/CD pipelines. Cloud Platforms: Deep experience developing and operating infrastructure across public cloud ecosystems, specifically AWS, Azure, and/or GCP. We value experience with enterprise platforms such as Salesforce or ServiceNow, as these skills translate well into our Enterprise-Grade Orchestration environment. What We Equip You With High-Impact Autonomy: A leadership environment where you will have real ownership over your team's direction, the latitude to make meaningful decisions, and participation in broader Engineering discussions. New Hire Orientation: A robust onboarding experience designed to integrate you smoothly into our technology, culture, and leadership model so you can show up for your team from day one. Continuous Enablement: Access to premier learning resources and dedicated learning time focused on both your continued technical growth and your evolution as an engineering leader. Sponsored Certifications: Full corporate sponsorship for professional technical certifications to advance your engineering credentials. The base salary range represents a good faith and reasonable estimate of the range at the time of posting. Actual compensation will be dependent on a number of factors including, but not limited to, the candidate's relevant work experience, qualifications, internal peer equity, and market and business conditions that exist when extending an offer. A discretionary bonus may be awarded in recognition of individual and company performance. In addition, Appian provides generous benefits offerings that include a 401(k) plan with company match, flexible time off, paid parental leave, medical, dental, and vision plans, life insurance, disability insurance, wellness programs, flexible spending accounts, health savings account contributions, an employee referral bonus program, and learning and development resources. Certain positions may be eligible for equity awards. Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation, commission, bonus, or benefit plans. Base Salary Range $175,000-$325,000 USD Tools and Resources Training and Development: During onboarding, we focus on equipping new hires with the skills and knowledge for success through department-specific training. Continuous learning is a central focus at Appian, with dedicated mentorship and the First-Friend program being widely utilized resources for new hires. Growth Opportunities: Appian provides a diverse array of growth and development opportunities, including our leadership program tailored for new and aspiring managers, a comprehensive library of specialized department training through Appian University, skills based training, and tuition reimbursement for those aiming to advance their education. This commitment ensures that employees have access to a holistic range of development opportunities. Community: We'll immerse you into our community rooted in respect starting on day one. Appian fosters inclusivity through our 8 employee-led affinity groups. These groups help employees build stronger internal and external networks by planning social, educational, and outreach activities to connect with Appianites and larger initiatives throughout the company. Benefits Appian offers a comprehensive benefits package designed to support your health, wellbeing, and financial future. Benefits may include health coverage, Employee Assistance Program (EAP) with free mental health support, life and disability insurance, an Employee Stock Purchase Program (ESPP), a retirement/pension plan, wellness dollars, tuition reimbursement, family-forming benefits and more. Benefits vary by country-please ask your Talent Acquisition contact for details specific to the location you are applying to. About Appian Appian provides AI automation for mission-critical work. We automate complex processes in large enterprises and governments . click apply for full job details
09/20/2026
Full time
Job Description Job Description Here at Appian, our values of Intensity and Excellence define who we are. We set high standards and live up to them, ensuring that everything we do is done with care and quality. We approach every challenge with ambition and commitment, holding ourselves and each other accountable to achieve the best results. When you join Appian, you'll be part of a passionate team dedicated to accomplishing hard things, together. Principal ML Platform Engineer Lead the engineering of software that matters - driving AI automation for the world's largest enterprises. This role is based at our h eadquarters in McLean, Virginia. Appian was built on a culture of in-person collaboration, which we believe is a key driver of our mission to be the best. Employees hired for this position are expected to be in the office 5 days a week to foster that culture and ensure we continue to thrive through shared ideas and teamwork. We believe being in the office provides more opportunities to come together and celebrate working with the exceptional people across Appian. About the Team Appian Engineering spans the full depth of our platform: from the foundational layers that power enterprise scale, to the AI capabilities redefining what automation can do. We operate in a highly collaborative, fast-paced environment focused on technical precision, continuous learning, and high code quality. By joining our team, you will solve real-world problems that directly shape how Appian delivers our AI-Powered Process Automation platform to enterprises around the world. The Opportunity As a Principal Software Engineer, you will serve as a technical linchpin for the team, bringing deep expertise in cloud-native architecture and a track record of influencing engineering direction beyond your immediate scope. Equipped with cutting-edge AI tooling, you will drive the design and delivery of high-complexity engineering solutions, set the technical bar for the team, and lead other engineers towards solutions of real complexity to ensure flawless Enterprise-Grade Orchestration. What You'll Do Develop Clean Software: Architect, build, and optimize high-performance software systems while maintaining a strong personal technical presence on the team. Lead Platform Modernization: Spearhead strategic technological changes and champion code refactoring efforts to keep the core Appian codebase cutting-edge, modern, and performant. Engineer with AI: Use AI coding tools fluently as a force multiplier: generating, reviewing, and critically evaluating AI-assisted code to ship faster without compromising quality or correctness. Lead Architecture & Delivery: Drive technical story breakdowns, acceptance criteria, and architectural design across complex, multi-tier application layers - from feature scoping through implementation. Optimize Performance & Scale: Manage product availability, latency, scalability, and efficiency by engineering deep reliability into our core software systems and performing advanced system tuning. Drive Engineering Excellence: Radiate development best practices across the department, perform meticulous code reviews on design and implementation, and build automation frameworks to prevent problem recurrence. Lead & Grow Engineers: Actively coach and mentor engineers at multiple levels, identify and close skill gaps on the team, and take ownership of accelerating the technical growth of those around you. Influence Technical Documentation: Share your expert domain knowledge regularly across the department, building a reputation as a vital resource and publishing high-quality content to Engineering's permanent documentation site. Required Qualifications Education: Minimum of a Bachelor of Science degree in Computer Science or a related technical/analytical discipline. (Equivalent experience is not accepted in lieu of a degree). Experience: 10+ years of relevant software development experience with a BS (or 8+ years of experience paired with a Master of Science in Computer Science or related field). Technical Mastery: Expert coding, scripting, and debugging proficiency in one or more core enterprise programming languages, specifically Java, Python, or Go. Domain Expertise: Deep working knowledge of distributed systems, cloud infrastructure, and the ability to contribute meaningfully at a senior individual contributor level within that space. Cross-Team Influence: Demonstrated ability to drive technical decisions and shape engineering practices beyond a single team or project scope. AI-Augmented Development: Demonstrated experience using AI coding assistants and a strong ability to evaluate, coach others on, and selectively apply AI-generated code in a production engineering context. Production Mastery: Proven experience developing, optimizing, and maintaining a high-volume, mission-critical production service environment. Communication & Alignment: Exceptional ability to communicate highly technical architectures verbally, visually, and in writing to diverse engineering audiences. Preferred Qualifications Cloud Architecture: Strong experience designing microservices, working with containerization (Docker, Kubernetes), and implementing modern CI/CD pipelines. Cloud Platforms: Deep experience developing and operating infrastructure across public cloud ecosystems, specifically AWS, Azure, and/or GCP. We value experience with enterprise platforms such as Salesforce or ServiceNow, as these skills translate well into our Enterprise-Grade Orchestration environment. What We Equip You With High-Impact Autonomy: A leadership environment where you will have real ownership over your team's direction, the latitude to make meaningful decisions, and participation in broader Engineering discussions. New Hire Orientation: A robust onboarding experience designed to integrate you smoothly into our technology, culture, and leadership model so you can show up for your team from day one. Continuous Enablement: Access to premier learning resources and dedicated learning time focused on both your continued technical growth and your evolution as an engineering leader. Sponsored Certifications: Full corporate sponsorship for professional technical certifications to advance your engineering credentials. The base salary range represents a good faith and reasonable estimate of the range at the time of posting. Actual compensation will be dependent on a number of factors including, but not limited to, the candidate's relevant work experience, qualifications, internal peer equity, and market and business conditions that exist when extending an offer. A discretionary bonus may be awarded in recognition of individual and company performance. In addition, Appian provides generous benefits offerings that include a 401(k) plan with company match, flexible time off, paid parental leave, medical, dental, and vision plans, life insurance, disability insurance, wellness programs, flexible spending accounts, health savings account contributions, an employee referral bonus program, and learning and development resources. Certain positions may be eligible for equity awards. Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation, commission, bonus, or benefit plans. Base Salary Range $175,000-$325,000 USD Tools and Resources Training and Development: During onboarding, we focus on equipping new hires with the skills and knowledge for success through department-specific training. Continuous learning is a central focus at Appian, with dedicated mentorship and the First-Friend program being widely utilized resources for new hires. Growth Opportunities: Appian provides a diverse array of growth and development opportunities, including our leadership program tailored for new and aspiring managers, a comprehensive library of specialized department training through Appian University, skills based training, and tuition reimbursement for those aiming to advance their education. This commitment ensures that employees have access to a holistic range of development opportunities. Community: We'll immerse you into our community rooted in respect starting on day one. Appian fosters inclusivity through our 8 employee-led affinity groups. These groups help employees build stronger internal and external networks by planning social, educational, and outreach activities to connect with Appianites and larger initiatives throughout the company. Benefits Appian offers a comprehensive benefits package designed to support your health, wellbeing, and financial future. Benefits may include health coverage, Employee Assistance Program (EAP) with free mental health support, life and disability insurance, an Employee Stock Purchase Program (ESPP), a retirement/pension plan, wellness dollars, tuition reimbursement, family-forming benefits and more. Benefits vary by country-please ask your Talent Acquisition contact for details specific to the location you are applying to. About Appian Appian provides AI automation for mission-critical work. We automate complex processes in large enterprises and governments . click apply for full job details
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work-work that changes the world-is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Despite the "Engineer" title, this isn't a typical back-end coding job. This is a pre-sales and architectural advisory role for someone who is part scientist, part consultant, and part communicator. At the Staff / Principal level, you are expected to be a hybrid heavy hitter - someone who can talk shop with PhD data scientists and then walk into a boardroom and explain the ROI to a CEO. We aren't looking for your average expert. We're building a team of T-shaped professionals with: Deep Technical Roots - You know PyTorch, Kubernetes, and GPU clusters (NVIDIA/AMD) inside and out, and you know when XGBoost is a better choice than the latest frontier model. Broad Business Acumen - You understand verticals, e.g. how a hospital's AI needs differ from a hedge fund's, and can navigate both conversations with credibility. Collaborative Autonomy - With a global remit and potential for 30-60% travel, you are in charge of driving results within diverse collaborations. We're building a small, global Applied AI team to work directly with enterprise customers and partners - helping them turn AI ambitions into production outcomes on Pure's platform. This team brings genuine research and production depth to customer-facing engagements: leading technical advisory sessions, shaping AI strategy for some of the world's largest companies, and contributing to Pure's growing reputation in AI through publications, conferences, and thought leadership. The team also ensures Pure's account teams can effectively qualify and position AI advisory opportunities across their territories, and provides enablement to our internal field and channel partner communities. The role requires working cross-functionally across Solutions, Marketing, Alliances, Engineering, Enablement, and Sales to align initiatives and go-to-market campaigns - all while maintaining genuine technical depth in enterprise AI. If this sounds like you, come join one of the most exciting teams at Pure. WHAT YOU'LL DO: Lead technical discovery and advisory engagements with customers to identify high-value AI use cases relevant to their industry and data Advise on end-to-end AI deployment - model selection, training/fine-tuning strategies, inference optimization, data pipeline design, and evaluation/alignment/safety frameworks - optimized for Pure's platform Collaborate cross-functionally to qualify and position AI opportunities, and develop proof-of-value prototypes that translate technical performance into business outcomes Create AI enablement content for field teams, partners, and customers - technical walkthroughs, qualification guides, workshops, and vertical-specific use case frameworks Present at industry conferences, publish technical content and peer-reviewed research, and represent Pure in engagements with strategic technology partners (NVIDIA, AMD, cloud providers, MSPs) Operate autonomously in ambiguous environments - independently scoping high-impact initiatives and driving them to completion at the right pace WHAT YOU BRING: Advanced degree in a quantitative field (Computer Science, Physics, Mathematics, Engineering, or related), or equivalent demonstrated through publications and production system experience 8+ years building and deploying AI/ML systems in cloud or on-prem environments Deep expertise in modern AI/ML - large language models, distributed training, inference optimization, agentic systems, evaluation/alignment frameworks, and classical ML Fluency with the modern AI stack (PyTorch, vLLM, Ray, Kubernetes) and data platforms (Spark, Snowflake, Kafka, etc.) Able to scope and carry out research projects that support critical business objectives, both independently and collaboratively Excellent written, verbal, and presentation skills - equally clear with hands-on data scientists and C-suite decision-makers WHAT SETS YOU APART: Experience with large-scale AI infrastructure: GPU clusters, high-performance storage, containerized deployments Experience in customer-facing technical advisory or consulting roles with enterprise accounts Experience leading independent, multi-year research from concept to publication or large-scale production deployment Exposure to multiple industry verticals (financial services, healthcare, telco, manufacturing) Publication record, conference presentations, or recognized technical presence Background combining applied research with shipping production systems Familiarity with CRM/opportunity management systems (Salesforce preferred) Salary ranges are determined based on role, level and location. For positions open to candidates in multiple geographical locations, the base salary range is reflective of the labor market across the applicable locations. This role may be eligible for incentive pay and/or equity. There is no application deadline and we accept applications on an ongoing basis until the job is filled. The annual base salary range is: $171,500 - $257,600 USD WHAT YOU CAN EXPECT FROM US: Innovation : We celebrate those who think critically, like a challenge, and aspire to be trailblazers. Growth : We give you the space and support to grow along with us and to contribute to something meaningful. We have been named Fortune's Best Workplaces in Technology , Fortune's Best Workplaces in the Bay Area , and certified as a Great Place to Work ! Team : We build each other up and set aside ego for the greater good. And because we understand the value of bringing your full and best self to work, we offer a variety of perks to manage a healthy balance, including flexible time off, wellness resources, and company-sponsored team events. Check out for more information. ACCOMMODATIONS AND ACCESSIBILITY: Candidates with disabilities may request accommodations for all aspects of our hiring process. For more on this, contact us at if you're invited to an interview. OUR COMMITMENT TO A STRONG AND INCLUSIVE TEAM: We're forging a future where everyone finds their rightful place and where every voice matters. Where uniqueness isn't just accepted but embraced. That's why we are committed to fostering the growth and development of every person, cultivating a sense of community through our Employee Resource Groups and advocating for inclusive leadership. Everpure is proud to be an equal opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or any other characteristic legally protected by the laws of the jurisdiction in which you are being considered for hire. Join us and bring your best. Bring your bold. Pure and simple.
09/20/2026
Full time
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work-work that changes the world-is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE Despite the "Engineer" title, this isn't a typical back-end coding job. This is a pre-sales and architectural advisory role for someone who is part scientist, part consultant, and part communicator. At the Staff / Principal level, you are expected to be a hybrid heavy hitter - someone who can talk shop with PhD data scientists and then walk into a boardroom and explain the ROI to a CEO. We aren't looking for your average expert. We're building a team of T-shaped professionals with: Deep Technical Roots - You know PyTorch, Kubernetes, and GPU clusters (NVIDIA/AMD) inside and out, and you know when XGBoost is a better choice than the latest frontier model. Broad Business Acumen - You understand verticals, e.g. how a hospital's AI needs differ from a hedge fund's, and can navigate both conversations with credibility. Collaborative Autonomy - With a global remit and potential for 30-60% travel, you are in charge of driving results within diverse collaborations. We're building a small, global Applied AI team to work directly with enterprise customers and partners - helping them turn AI ambitions into production outcomes on Pure's platform. This team brings genuine research and production depth to customer-facing engagements: leading technical advisory sessions, shaping AI strategy for some of the world's largest companies, and contributing to Pure's growing reputation in AI through publications, conferences, and thought leadership. The team also ensures Pure's account teams can effectively qualify and position AI advisory opportunities across their territories, and provides enablement to our internal field and channel partner communities. The role requires working cross-functionally across Solutions, Marketing, Alliances, Engineering, Enablement, and Sales to align initiatives and go-to-market campaigns - all while maintaining genuine technical depth in enterprise AI. If this sounds like you, come join one of the most exciting teams at Pure. WHAT YOU'LL DO: Lead technical discovery and advisory engagements with customers to identify high-value AI use cases relevant to their industry and data Advise on end-to-end AI deployment - model selection, training/fine-tuning strategies, inference optimization, data pipeline design, and evaluation/alignment/safety frameworks - optimized for Pure's platform Collaborate cross-functionally to qualify and position AI opportunities, and develop proof-of-value prototypes that translate technical performance into business outcomes Create AI enablement content for field teams, partners, and customers - technical walkthroughs, qualification guides, workshops, and vertical-specific use case frameworks Present at industry conferences, publish technical content and peer-reviewed research, and represent Pure in engagements with strategic technology partners (NVIDIA, AMD, cloud providers, MSPs) Operate autonomously in ambiguous environments - independently scoping high-impact initiatives and driving them to completion at the right pace WHAT YOU BRING: Advanced degree in a quantitative field (Computer Science, Physics, Mathematics, Engineering, or related), or equivalent demonstrated through publications and production system experience 8+ years building and deploying AI/ML systems in cloud or on-prem environments Deep expertise in modern AI/ML - large language models, distributed training, inference optimization, agentic systems, evaluation/alignment frameworks, and classical ML Fluency with the modern AI stack (PyTorch, vLLM, Ray, Kubernetes) and data platforms (Spark, Snowflake, Kafka, etc.) Able to scope and carry out research projects that support critical business objectives, both independently and collaboratively Excellent written, verbal, and presentation skills - equally clear with hands-on data scientists and C-suite decision-makers WHAT SETS YOU APART: Experience with large-scale AI infrastructure: GPU clusters, high-performance storage, containerized deployments Experience in customer-facing technical advisory or consulting roles with enterprise accounts Experience leading independent, multi-year research from concept to publication or large-scale production deployment Exposure to multiple industry verticals (financial services, healthcare, telco, manufacturing) Publication record, conference presentations, or recognized technical presence Background combining applied research with shipping production systems Familiarity with CRM/opportunity management systems (Salesforce preferred) Salary ranges are determined based on role, level and location. For positions open to candidates in multiple geographical locations, the base salary range is reflective of the labor market across the applicable locations. This role may be eligible for incentive pay and/or equity. There is no application deadline and we accept applications on an ongoing basis until the job is filled. The annual base salary range is: $171,500 - $257,600 USD WHAT YOU CAN EXPECT FROM US: Innovation : We celebrate those who think critically, like a challenge, and aspire to be trailblazers. Growth : We give you the space and support to grow along with us and to contribute to something meaningful. We have been named Fortune's Best Workplaces in Technology , Fortune's Best Workplaces in the Bay Area , and certified as a Great Place to Work ! Team : We build each other up and set aside ego for the greater good. And because we understand the value of bringing your full and best self to work, we offer a variety of perks to manage a healthy balance, including flexible time off, wellness resources, and company-sponsored team events. Check out for more information. ACCOMMODATIONS AND ACCESSIBILITY: Candidates with disabilities may request accommodations for all aspects of our hiring process. For more on this, contact us at if you're invited to an interview. OUR COMMITMENT TO A STRONG AND INCLUSIVE TEAM: We're forging a future where everyone finds their rightful place and where every voice matters. Where uniqueness isn't just accepted but embraced. That's why we are committed to fostering the growth and development of every person, cultivating a sense of community through our Employee Resource Groups and advocating for inclusive leadership. Everpure is proud to be an equal opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or any other characteristic legally protected by the laws of the jurisdiction in which you are being considered for hire. Join us and bring your best. Bring your bold. Pure and simple.
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work-work that changes the world-is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a key contributor on our System Hardware Development team, you will drive the end-to-end architecture, validation, and quality of enterprise-class storage hardware systems that power mission-critical data environments. You will bridge complex engineering concepts with real-world deployment by leading cross-functional alignment across internal engineering, Joint Development Manufacturing (JDM) partners, operations, and support teams. In this role, your technical leadership ensures unmatched system reliability, performance, and scalability across our next-generation storage portfolio while establishing design methodologies that accelerate product innovation. WHAT YOU'LL DO Drive Next-Generation Hardware Architecture & Design : Lead schematic capture, high-speed PCB layout, signal integrity analysis, and x86 system-level co-development alongside JDM partners to deliver scalable, high-performance enterprise storage platforms from concept through production. Architect Comprehensive Qualification & Automation Frameworks : Oversee functional specifications and automated stress testing environments-leveraging Linux, Python, PCIe, DDR, Ethernet, and serial bus protocols-to guarantee sub-system reliability and accelerate time-to-market. Champion Cross-Functional Hardware Operations & System Hardening : Partner with Manufacturing, Operations, and Escalation teams to streamline production handoffs for new storage sub-systems and perform deep root-cause failure analysis that permanently hardens enterprise products. Establish Quality-by-Design Standards & Best Practices : Define hardware validation methodologies, collaborate across organizational boundaries, and build engineering networks that continuously elevate quality benchmarks across all product development lifecycles. WHAT YOU BRING Enterprise Hardware & x86 Architecture Mastery : Expertise in x86 system architecture, high-speed PCB design layout, signal integrity analysis, schematic capture, and major bus technologies (PCIe, DDR, Ethernet, I2C, SPI) to independently validate enterprise hardware at electrical, protocol, and functional levels. Advanced System Testing & Automation : Ability to design and automate system-level stress tests (using Linux environments, Python/BASH scripting, PTU, LinPack, and IOmeter) and utilize high-speed oscilloscopes and bus analyzers to perform complex signal and protocol trace analysis. Cross-Functional Leadership & Complex Problem Solving : Capability to lead root-cause failure analysis on complex technical issues, establish engineering best practices, and collaborate effectively with internal teams and external manufacturing partners to deliver production-ready systems. Hardware Logic & Diagnostic Fundamentals : Solid understanding of Verilog and CPLD design concepts combined with a rigorous, quality-focused mindset for hardware design and system-level troubleshooting. We are primarily an in-office environment and therefore, you will be expected to work from the Santa Clara, CA office in compliance with Everpure's policies, unless you are on PTO, or work travel, or other approved leave. Salary ranges are determined based on role, level and location. For positions open to candidates in multiple geographical locations, the base salary range is reflective of the labor market across the applicable locations. This role may be eligible for incentive pay and/or equity. There is no application deadline and we accept applications on an ongoing basis until the job is filled. The annual base salary range is: $217,000 - $359,000 USD WHAT YOU CAN EXPECT FROM US: Innovation : We celebrate those who think critically, like a challenge, and aspire to be trailblazers. Growth : We give you the space and support to grow along with us and to contribute to something meaningful. We have been named Fortune's Best Workplaces in Technology , Fortune's Best Workplaces in the Bay Area , and certified as a Great Place to Work ! Team : We build each other up and set aside ego for the greater good. And because we understand the value of bringing your full and best self to work, we offer a variety of perks to manage a healthy balance, including flexible time off, wellness resources, and company-sponsored team events. Check out for more information. ACCOMMODATIONS AND ACCESSIBILITY: Candidates with disabilities may request accommodations for all aspects of our hiring process. For more on this, contact us at if you're invited to an interview. OUR COMMITMENT TO A STRONG AND INCLUSIVE TEAM: We're forging a future where everyone finds their rightful place and where every voice matters. Where uniqueness isn't just accepted but embraced. That's why we are committed to fostering the growth and development of every person, cultivating a sense of community through our Employee Resource Groups and advocating for inclusive leadership. Everpure is proud to be an equal opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or any other characteristic legally protected by the laws of the jurisdiction in which you are being considered for hire. Join us and bring your best. Bring your bold. Pure and simple.
09/20/2026
Full time
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work-work that changes the world-is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE As a key contributor on our System Hardware Development team, you will drive the end-to-end architecture, validation, and quality of enterprise-class storage hardware systems that power mission-critical data environments. You will bridge complex engineering concepts with real-world deployment by leading cross-functional alignment across internal engineering, Joint Development Manufacturing (JDM) partners, operations, and support teams. In this role, your technical leadership ensures unmatched system reliability, performance, and scalability across our next-generation storage portfolio while establishing design methodologies that accelerate product innovation. WHAT YOU'LL DO Drive Next-Generation Hardware Architecture & Design : Lead schematic capture, high-speed PCB layout, signal integrity analysis, and x86 system-level co-development alongside JDM partners to deliver scalable, high-performance enterprise storage platforms from concept through production. Architect Comprehensive Qualification & Automation Frameworks : Oversee functional specifications and automated stress testing environments-leveraging Linux, Python, PCIe, DDR, Ethernet, and serial bus protocols-to guarantee sub-system reliability and accelerate time-to-market. Champion Cross-Functional Hardware Operations & System Hardening : Partner with Manufacturing, Operations, and Escalation teams to streamline production handoffs for new storage sub-systems and perform deep root-cause failure analysis that permanently hardens enterprise products. Establish Quality-by-Design Standards & Best Practices : Define hardware validation methodologies, collaborate across organizational boundaries, and build engineering networks that continuously elevate quality benchmarks across all product development lifecycles. WHAT YOU BRING Enterprise Hardware & x86 Architecture Mastery : Expertise in x86 system architecture, high-speed PCB design layout, signal integrity analysis, schematic capture, and major bus technologies (PCIe, DDR, Ethernet, I2C, SPI) to independently validate enterprise hardware at electrical, protocol, and functional levels. Advanced System Testing & Automation : Ability to design and automate system-level stress tests (using Linux environments, Python/BASH scripting, PTU, LinPack, and IOmeter) and utilize high-speed oscilloscopes and bus analyzers to perform complex signal and protocol trace analysis. Cross-Functional Leadership & Complex Problem Solving : Capability to lead root-cause failure analysis on complex technical issues, establish engineering best practices, and collaborate effectively with internal teams and external manufacturing partners to deliver production-ready systems. Hardware Logic & Diagnostic Fundamentals : Solid understanding of Verilog and CPLD design concepts combined with a rigorous, quality-focused mindset for hardware design and system-level troubleshooting. We are primarily an in-office environment and therefore, you will be expected to work from the Santa Clara, CA office in compliance with Everpure's policies, unless you are on PTO, or work travel, or other approved leave. Salary ranges are determined based on role, level and location. For positions open to candidates in multiple geographical locations, the base salary range is reflective of the labor market across the applicable locations. This role may be eligible for incentive pay and/or equity. There is no application deadline and we accept applications on an ongoing basis until the job is filled. The annual base salary range is: $217,000 - $359,000 USD WHAT YOU CAN EXPECT FROM US: Innovation : We celebrate those who think critically, like a challenge, and aspire to be trailblazers. Growth : We give you the space and support to grow along with us and to contribute to something meaningful. We have been named Fortune's Best Workplaces in Technology , Fortune's Best Workplaces in the Bay Area , and certified as a Great Place to Work ! Team : We build each other up and set aside ego for the greater good. And because we understand the value of bringing your full and best self to work, we offer a variety of perks to manage a healthy balance, including flexible time off, wellness resources, and company-sponsored team events. Check out for more information. ACCOMMODATIONS AND ACCESSIBILITY: Candidates with disabilities may request accommodations for all aspects of our hiring process. For more on this, contact us at if you're invited to an interview. OUR COMMITMENT TO A STRONG AND INCLUSIVE TEAM: We're forging a future where everyone finds their rightful place and where every voice matters. Where uniqueness isn't just accepted but embraced. That's why we are committed to fostering the growth and development of every person, cultivating a sense of community through our Employee Resource Groups and advocating for inclusive leadership. Everpure is proud to be an equal opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or any other characteristic legally protected by the laws of the jurisdiction in which you are being considered for hire. Join us and bring your best. Bring your bold. Pure and simple.
Every token a large language model generates depends on data reaching the right accelerator at the right moment. As AI models outgrow any single chip, the network between accelerators becomes the bottleneck that decides how fast - and how affordably - the world's largest models can serve real users. That network layer is what our team builds. We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving into separate prefill and decode pools and moving the model's KV cache between them at the limit of what the hardware allows. Get it right and users get answers in milliseconds; get it wrong and the fastest accelerators in the world sit idle waiting on data. You'll build components of the high-speed transfer path that make that difference, and you'll learn to measure success in how close we run to the theoretical peak of the machine. In this role you will: Build and optimize the low-level data-movement software that transfers KV cache and activations across accelerators, servers, and heterogeneous memory - over AWS's highest-performance network fabric. Profile real workloads, find the true bottleneck, and close the gap between "it works" and "it runs fast" - pushing components toward the hardware's limit. Work across the stack - from network transport up to the inference frameworks - learning from the teams building the chips, runtime, and models. Deliver features that ship to our largest clusters, for our largest customers, serving the largest AI models in production. What we're looking for: Strong C/C++ and a genuine interest in low-level, performance-critical systems - solid command of Linux, memory, and writing fast code. The instinct to ask "how fast could this go?" and the discipline to measure it. Exposure to high-speed networking, HPC interconnects, or GPU/accelerator systems (RDMA, InfiniBand, libfabric, UCX, NCCL, MPI) is a strong plus; embedded-systems experience is welcome. Prior AI/ML experience is not required - if you're a strong systems engineer eager to learn, we'll teach you the ML side. If you like solving genuinely hard problems, working alongside HPC and ML customers, iterating fast, and shipping at a scale few places can offer, come join us. You'll work alongside senior engineers and Principal Engineers who've built this layer from the ground up, with real room to grow your scope and technical depth - on a team at the leading edge of AI/ML infrastructure. About the team: You'd be joining Annapurna Labs, an integral part of AWS. Annapurna designs the hardware and software building blocks behind EC2 - every EC2 instance runs on hardware we designed. We specialize in the chips, systems, and software that optimize the AWS customer experience, and this team sits where AI meets the silicon and the network underneath it. A day in the life Annapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS. We have mixed discipline orgs, you'd be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you'll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Experience with C/C++ PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
Every token a large language model generates depends on data reaching the right accelerator at the right moment. As AI models outgrow any single chip, the network between accelerators becomes the bottleneck that decides how fast - and how affordably - the world's largest models can serve real users. That network layer is what our team builds. We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving into separate prefill and decode pools and moving the model's KV cache between them at the limit of what the hardware allows. Get it right and users get answers in milliseconds; get it wrong and the fastest accelerators in the world sit idle waiting on data. You'll build components of the high-speed transfer path that make that difference, and you'll learn to measure success in how close we run to the theoretical peak of the machine. In this role you will: Build and optimize the low-level data-movement software that transfers KV cache and activations across accelerators, servers, and heterogeneous memory - over AWS's highest-performance network fabric. Profile real workloads, find the true bottleneck, and close the gap between "it works" and "it runs fast" - pushing components toward the hardware's limit. Work across the stack - from network transport up to the inference frameworks - learning from the teams building the chips, runtime, and models. Deliver features that ship to our largest clusters, for our largest customers, serving the largest AI models in production. What we're looking for: Strong C/C++ and a genuine interest in low-level, performance-critical systems - solid command of Linux, memory, and writing fast code. The instinct to ask "how fast could this go?" and the discipline to measure it. Exposure to high-speed networking, HPC interconnects, or GPU/accelerator systems (RDMA, InfiniBand, libfabric, UCX, NCCL, MPI) is a strong plus; embedded-systems experience is welcome. Prior AI/ML experience is not required - if you're a strong systems engineer eager to learn, we'll teach you the ML side. If you like solving genuinely hard problems, working alongside HPC and ML customers, iterating fast, and shipping at a scale few places can offer, come join us. You'll work alongside senior engineers and Principal Engineers who've built this layer from the ground up, with real room to grow your scope and technical depth - on a team at the leading edge of AI/ML infrastructure. About the team: You'd be joining Annapurna Labs, an integral part of AWS. Annapurna designs the hardware and software building blocks behind EC2 - every EC2 instance runs on hardware we designed. We specialize in the chips, systems, and software that optimize the AWS customer experience, and this team sits where AI meets the silicon and the network underneath it. A day in the life Annapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS. We have mixed discipline orgs, you'd be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you'll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Experience with C/C++ PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
Amazon Development Center U.S., Inc.
Seattle, Washington
Amazon Quick is building the future of AI-powered work - an intelligent companion that connects to your tools, learns your context, and automates your workflows through autonomous agents. We're looking for a Principal Product Manager - Technical to own and drive our end-to-end user experiences - defining how millions of knowledge workers interact with AI agents across Web, Desktop, and Mobile, and shaping the surfaces where autonomous work becomes visible, trustworthy, and delightful. You will own the vision for Amazon Quick's Web, Desktop, Mobile, and Extensions experiences - the primary surfaces through which customers discover, interact with, and benefit from AI-powered agents. You'll define how conversation, collaboration, and autonomous task execution come together in a unified experience that adapts to how people actually work: in the browser, at their desk, on the go, and everywhere in between. You'll drive the strategy for how agents surface insights, take action, and earn user trust through thoughtful UX that balances power with simplicity - ensuring Amazon Quick becomes the indispensable AI companion that workers reach for first, on every device. Key job responsibilities • Own the Amazon Quick experience end-to-end across surfaces- define the vision, roadmap, and success metrics for how customers interact with AI agents on every surface. • Drive data-informed product decisions through funnel instrumentation, experimentation, and cohort analysis. Prioritize complex deliverables across engineering, design, and science teams while managing technical trade-offs. • Define the unified experience framework - interaction patterns, navigation models, and design systems - that deliver a cohesive, high-quality experience across all platforms while leveraging each surface's unique strengths. • Own cross-platform orchestration - how the experience adapts across form factors, handles context continuity between devices, and delivers seamless transitions from desktop deep-work to mobile triage and back. • Drive the customer adoption journey from first-run onboarding through power-user engagement, ensuring every surface earns habitual daily use. • Scale the experience to meet organizational goals. Lead cross-functional execution with Engineering, UX Design, Applied Science, and GTM teams to ship world-class products across all platforms. BASIC QUALIFICATIONS - 7+ years of working as a Technical Product Manager experience - 5+ years of technical (software development, network development, IT, other related) experience - Experience delivering large-scale SaaS, PaaS or LaaS products where you are responsible for the full product lifecycle, from concept through GTM (go to market) PREFERRED QUALIFICATIONS - Experience designing adaptive systems that vary behavior by device, context, or user state - Track record of decomposing ambiguous problems into structured frameworks - Direct collaboration with applied scientists or ML engineers - co-defining inputs/outputs, evaluation criteria, or routing logic - Comfort with probabilistic/LLM-based systems - can reason about capabilities, failure modes, and offline evaluation without needing to build them Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, WA, Seattle - 181 000.00 USD annually
09/20/2026
Full time
Amazon Quick is building the future of AI-powered work - an intelligent companion that connects to your tools, learns your context, and automates your workflows through autonomous agents. We're looking for a Principal Product Manager - Technical to own and drive our end-to-end user experiences - defining how millions of knowledge workers interact with AI agents across Web, Desktop, and Mobile, and shaping the surfaces where autonomous work becomes visible, trustworthy, and delightful. You will own the vision for Amazon Quick's Web, Desktop, Mobile, and Extensions experiences - the primary surfaces through which customers discover, interact with, and benefit from AI-powered agents. You'll define how conversation, collaboration, and autonomous task execution come together in a unified experience that adapts to how people actually work: in the browser, at their desk, on the go, and everywhere in between. You'll drive the strategy for how agents surface insights, take action, and earn user trust through thoughtful UX that balances power with simplicity - ensuring Amazon Quick becomes the indispensable AI companion that workers reach for first, on every device. Key job responsibilities • Own the Amazon Quick experience end-to-end across surfaces- define the vision, roadmap, and success metrics for how customers interact with AI agents on every surface. • Drive data-informed product decisions through funnel instrumentation, experimentation, and cohort analysis. Prioritize complex deliverables across engineering, design, and science teams while managing technical trade-offs. • Define the unified experience framework - interaction patterns, navigation models, and design systems - that deliver a cohesive, high-quality experience across all platforms while leveraging each surface's unique strengths. • Own cross-platform orchestration - how the experience adapts across form factors, handles context continuity between devices, and delivers seamless transitions from desktop deep-work to mobile triage and back. • Drive the customer adoption journey from first-run onboarding through power-user engagement, ensuring every surface earns habitual daily use. • Scale the experience to meet organizational goals. Lead cross-functional execution with Engineering, UX Design, Applied Science, and GTM teams to ship world-class products across all platforms. BASIC QUALIFICATIONS - 7+ years of working as a Technical Product Manager experience - 5+ years of technical (software development, network development, IT, other related) experience - Experience delivering large-scale SaaS, PaaS or LaaS products where you are responsible for the full product lifecycle, from concept through GTM (go to market) PREFERRED QUALIFICATIONS - Experience designing adaptive systems that vary behavior by device, context, or user state - Track record of decomposing ambiguous problems into structured frameworks - Direct collaboration with applied scientists or ML engineers - co-defining inputs/outputs, evaluation criteria, or routing logic - Comfort with probabilistic/LLM-based systems - can reason about capabilities, failure modes, and offline evaluation without needing to build them Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, WA, Seattle - 181 000.00 USD annually
American Honda Motor Co., Inc. seeks a Principal Applications Engineer to lead the design, development, and optimization of critical applications supporting manufacturing, supply chain, and enterprise operations. In this innovation-driven role, you will define technical architectures, integrate with manufacturing and ERP systems, and ensure reliability, performance, and security. You will collaborate with cross-functional teams, mentor engineers, and champion best practices, enabling scalable, high-quality solutions that advance safety, sustainability, and continuous improvement across Honda's U.S. operations. Responsibilities Lead design, development, and optimization of manufacturing and enterprise applications supporting Honda operations Collaborate with cross-functional engineering, manufacturing, and IT teams to define technical requirements and architectures Own end-to-end application lifecycle including design, coding standards, testing, deployment, and performance tuning Integrate applications with manufacturing systems, ERP, and data platforms to improve efficiency and quality Drive technical roadmap, best practices, and governance for application platforms Mentor and guide engineers, reviewing designs and code for scalability, reliability, and security Troubleshoot complex production issues and implement robust, sustainable fixes Contribute to innovation initiatives in automation, analytics, and Industry 4.0 solutions Required Skills Enterprise application architecture Manufacturing/industrial systems integration C#/. NET or Java development RESTful APIs and microservices SQL and relational database design Cloud platforms (AWS/Azure/GCP) CI/CD pipelines and Dev Ops tools Application performance tuning and monitoring System design and UML modeling Secure coding and application security
09/20/2026
Full time
American Honda Motor Co., Inc. seeks a Principal Applications Engineer to lead the design, development, and optimization of critical applications supporting manufacturing, supply chain, and enterprise operations. In this innovation-driven role, you will define technical architectures, integrate with manufacturing and ERP systems, and ensure reliability, performance, and security. You will collaborate with cross-functional teams, mentor engineers, and champion best practices, enabling scalable, high-quality solutions that advance safety, sustainability, and continuous improvement across Honda's U.S. operations. Responsibilities Lead design, development, and optimization of manufacturing and enterprise applications supporting Honda operations Collaborate with cross-functional engineering, manufacturing, and IT teams to define technical requirements and architectures Own end-to-end application lifecycle including design, coding standards, testing, deployment, and performance tuning Integrate applications with manufacturing systems, ERP, and data platforms to improve efficiency and quality Drive technical roadmap, best practices, and governance for application platforms Mentor and guide engineers, reviewing designs and code for scalability, reliability, and security Troubleshoot complex production issues and implement robust, sustainable fixes Contribute to innovation initiatives in automation, analytics, and Industry 4.0 solutions Required Skills Enterprise application architecture Manufacturing/industrial systems integration C#/. NET or Java development RESTful APIs and microservices SQL and relational database design Cloud platforms (AWS/Azure/GCP) CI/CD pipelines and Dev Ops tools Application performance tuning and monitoring System design and UML modeling Secure coding and application security
Who We Are As the largest private-sector power producer in the world and the nation's largest producer of clean and reliable energy, Constellation is focused on our purpose: lighting the way to a brilliant tomorrow for all. We have been the leader in clean energy production for more than a decade, and we are cultivating a workplace where our employees can grow, thrive, and contribute. Now integrated with Calpine, our portfolio includes 55 gigawatts of capacity from nuclear, natural gas, geothermal, hydro, wind and solar facilities, with the generating capacity to power the equivalent of 27 million homes. Our culture and employee experience make it clear: We are powered by passion and purpose. Together, we're creating healthier communities and a cleaner planet, and our people are the driving force behind our success. At Constellation, you can build a fulfilling career with opportunities to learn, grow and make an impact. By doing our best work and meeting new challenges, we can accomplish great things. Join us in meeting the country's energy needs today and tomorrow. Total Rewards Constellation offers an extensive selection of benefits and rewards to help our employees thrive professionally and personally. We provide competitive compensation and a wide-range of benefits that support both employees and their families, helping them prepare for the future. In addition to highly competitive salaries, eligible employees are offered a bonus program, 401(k) with company match, employee stock purchase program comprehensive medical, dental and vision benefits, including robust wellbeing programs disability and life insurance benefits paid time off for vacation, holidays, and sick days and much more. Expected salary range of $172,800 to $192,000, varies based on experience, along with comprehensive benefits package that includes bonus and 401(k). Primary Purpose of Position The Principal IT Architect is a cross-project and cross-discipline role that is charged with the creation, governance, maintenance and communication of Constellation's current-state and future-state Architectures. This person will work with the Constellation leadership to set direction for the Architecture organization and support the IT/Business functions. The Principal IT Architect is responsible for providing infrastructure/platform/governance strategy, architecture, and technical leadership to serve a portfolio of technologies that acts as a foundation for Constellation IT business applications, services, and solutions. Principal architects must have a broad base of technical and business understanding as well as the attention to detail to transform a strategy into an actionable plan. The successful Principal IT Architect will demonstrate knowledge to navigate through recent technology trends and connect it with the business goals to act as an influencer. Key focus area: The Responsible AI (RAI) Champion serves as a critical advocate for ethical and responsible AI implementation within the organization. They bridge the gap between technical excellence and ethical considerations, ensuring that AI solutions align with organizational values and societal impact. The Responsible AI Champion will be a member of a larger Responsible AI team that exists to execute governance as necessary to facilitate the following functions: Operate in accordance with the Constellation RAI Policy and RAI Risk Policy (including Risk Matrix) Approvals and dispositions on net new or material modifications to existing applications of AI to business solutions. Catalog all AI use cases (approved, implemented, or denied) as part of a backlog for prioritization. Inform, advise, and assist in maintaining compliance with RAI policies, laws, regulations, Constellation RAI policies (Corporate & Risk), technical standards, and best practices. Set and catalog approved and unapproved architectural patterns of RAI technology. Advocate for and support democratization and adoption of RAI capabilities and infuse them in established functions within Constellation Support continuous improvement and RAI program maturity by monitoring AI current events, industry/non-industry benchmarking, legal and regulatory policy changes. To be successful in this role, The Principal Architect must collaborate with Business and IT leadership, application portfolio owners, infrastructure architects, solution architects, enterprise architects, relevant IT engineering /operations teams and consulting partners, among many others. This role must be proactive, open-minded, have a curious mind and a questioning attitude to drive continuous improvement. Strong written, verbal, and interpersonal skills are required. Deep experience delivering critical strategic direction and design artifacts as well as presenting these deliverables Primary Duties and Accountabilities Provide technical and security expertise to IT and business teams to identify security technology solutions and develop security reference architectures and strategies to achieve business results. Manage and monitor multiple intake channels for evaluating proposed and active implementations of AI that will vary in stages of conception (Ideation, Analysis, Design, Development). Ensure appropriate implementation of technology within both the development and production environments by leveraging the Responsible AI processes. Provide technological expertise and advice to IT and Business leadership in the development of strategic information technology plans to support business strategies tied to AI. Establish, maintain, and enhance relationships with business and IT partners. Communicate status to key stakeholders on a regular basis through partnership with the Responsible AI program lead. Participate in governance mechanisms and ensure that architecture deliverables exceed AI governance standards. Maintain awareness of trends and issues in area of technical expertise evaluate new technologies or technology opportunities and provide analysis of their potential impact to the business. Provide coaching/ mentorship for IT personnel. Participate in career development and recognition activities. Proven oral and written communication and presentation skills. Promote inclusion and foster teamwork, collaboration, and a learning organization. Detail-oriented with excellent analytical and problem-solving skills. Minimum Qualifications Bachelor's degree in Information Systems, Computer Science, or Business Administration 10-years technical experience designing, developing, and delivering medium to large scale software solutions. 5-years technical experience designing, developing, and delivering AI/ML solutions. Experience utilizing Microsoft technologies and Azure components Recognized subject matter expert in Generative AI, Machine Learning, or Data Science Knowledge of cybersecurity fundamentals Ability to work independently, collaboratively with minimal supervision Perseverance in the face of setbacks Must be willing to travel 25% Preferred Qualifications Master's Degree or MBA Professional IT Architecture education or certification Demonstrated experience contributing to Responsible AI governance initiatives
09/20/2026
Full time
Who We Are As the largest private-sector power producer in the world and the nation's largest producer of clean and reliable energy, Constellation is focused on our purpose: lighting the way to a brilliant tomorrow for all. We have been the leader in clean energy production for more than a decade, and we are cultivating a workplace where our employees can grow, thrive, and contribute. Now integrated with Calpine, our portfolio includes 55 gigawatts of capacity from nuclear, natural gas, geothermal, hydro, wind and solar facilities, with the generating capacity to power the equivalent of 27 million homes. Our culture and employee experience make it clear: We are powered by passion and purpose. Together, we're creating healthier communities and a cleaner planet, and our people are the driving force behind our success. At Constellation, you can build a fulfilling career with opportunities to learn, grow and make an impact. By doing our best work and meeting new challenges, we can accomplish great things. Join us in meeting the country's energy needs today and tomorrow. Total Rewards Constellation offers an extensive selection of benefits and rewards to help our employees thrive professionally and personally. We provide competitive compensation and a wide-range of benefits that support both employees and their families, helping them prepare for the future. In addition to highly competitive salaries, eligible employees are offered a bonus program, 401(k) with company match, employee stock purchase program comprehensive medical, dental and vision benefits, including robust wellbeing programs disability and life insurance benefits paid time off for vacation, holidays, and sick days and much more. Expected salary range of $172,800 to $192,000, varies based on experience, along with comprehensive benefits package that includes bonus and 401(k). Primary Purpose of Position The Principal IT Architect is a cross-project and cross-discipline role that is charged with the creation, governance, maintenance and communication of Constellation's current-state and future-state Architectures. This person will work with the Constellation leadership to set direction for the Architecture organization and support the IT/Business functions. The Principal IT Architect is responsible for providing infrastructure/platform/governance strategy, architecture, and technical leadership to serve a portfolio of technologies that acts as a foundation for Constellation IT business applications, services, and solutions. Principal architects must have a broad base of technical and business understanding as well as the attention to detail to transform a strategy into an actionable plan. The successful Principal IT Architect will demonstrate knowledge to navigate through recent technology trends and connect it with the business goals to act as an influencer. Key focus area: The Responsible AI (RAI) Champion serves as a critical advocate for ethical and responsible AI implementation within the organization. They bridge the gap between technical excellence and ethical considerations, ensuring that AI solutions align with organizational values and societal impact. The Responsible AI Champion will be a member of a larger Responsible AI team that exists to execute governance as necessary to facilitate the following functions: Operate in accordance with the Constellation RAI Policy and RAI Risk Policy (including Risk Matrix) Approvals and dispositions on net new or material modifications to existing applications of AI to business solutions. Catalog all AI use cases (approved, implemented, or denied) as part of a backlog for prioritization. Inform, advise, and assist in maintaining compliance with RAI policies, laws, regulations, Constellation RAI policies (Corporate & Risk), technical standards, and best practices. Set and catalog approved and unapproved architectural patterns of RAI technology. Advocate for and support democratization and adoption of RAI capabilities and infuse them in established functions within Constellation Support continuous improvement and RAI program maturity by monitoring AI current events, industry/non-industry benchmarking, legal and regulatory policy changes. To be successful in this role, The Principal Architect must collaborate with Business and IT leadership, application portfolio owners, infrastructure architects, solution architects, enterprise architects, relevant IT engineering /operations teams and consulting partners, among many others. This role must be proactive, open-minded, have a curious mind and a questioning attitude to drive continuous improvement. Strong written, verbal, and interpersonal skills are required. Deep experience delivering critical strategic direction and design artifacts as well as presenting these deliverables Primary Duties and Accountabilities Provide technical and security expertise to IT and business teams to identify security technology solutions and develop security reference architectures and strategies to achieve business results. Manage and monitor multiple intake channels for evaluating proposed and active implementations of AI that will vary in stages of conception (Ideation, Analysis, Design, Development). Ensure appropriate implementation of technology within both the development and production environments by leveraging the Responsible AI processes. Provide technological expertise and advice to IT and Business leadership in the development of strategic information technology plans to support business strategies tied to AI. Establish, maintain, and enhance relationships with business and IT partners. Communicate status to key stakeholders on a regular basis through partnership with the Responsible AI program lead. Participate in governance mechanisms and ensure that architecture deliverables exceed AI governance standards. Maintain awareness of trends and issues in area of technical expertise evaluate new technologies or technology opportunities and provide analysis of their potential impact to the business. Provide coaching/ mentorship for IT personnel. Participate in career development and recognition activities. Proven oral and written communication and presentation skills. Promote inclusion and foster teamwork, collaboration, and a learning organization. Detail-oriented with excellent analytical and problem-solving skills. Minimum Qualifications Bachelor's degree in Information Systems, Computer Science, or Business Administration 10-years technical experience designing, developing, and delivering medium to large scale software solutions. 5-years technical experience designing, developing, and delivering AI/ML solutions. Experience utilizing Microsoft technologies and Azure components Recognized subject matter expert in Generative AI, Machine Learning, or Data Science Knowledge of cybersecurity fundamentals Ability to work independently, collaboratively with minimal supervision Perseverance in the face of setbacks Must be willing to travel 25% Preferred Qualifications Master's Degree or MBA Professional IT Architecture education or certification Demonstrated experience contributing to Responsible AI governance initiatives
Job Description Job Description Our mission is to create the Experience of a Lifetime for our employees, so they can, in turn, create the Experience of a Lifetime for our guests. We own and operate the most renowned destination resorts in the world as well as regional and local ski areas outside major cities, and connect them all through one unrivaled network. We are looking for ambitious leaders, innovators and creators to join our talented team. If you're ready to pursue your fullest potential, we want to get to know you! Candidates for year-round positions are reviewed on a rolling basis. Applications will be accepted up to 90 days after the posting date, or until the position is filled (whichever is first). Job Summary: We are looking for a curious, driven, innovative machine learning engineer who takes initiative to solve problems and create environments that accelerate the development, deployment, and usage of data science models and AI to drive greater organizational impact. The Data Science & Data Engineering team within the Enterprise Analytics organization builds data assets, predictive models, analytical applications, and platforms across the organization. Our team collaborates with business stakeholders, analysts, and technology teams to tackle high-impact use cases with state-of-the-art models and tools to grow the business, streamline costs, and improve guest experiences. Job Specifications: Starting Wage: $140,000 - $185,000 + Annual Bonus Employment Type: Year Round Shift Type: Full Time hours Minimum Age: At least 18 years of age Housing Availability: No Job Responsibilities: Productionize ML models developed by data science into reliable, monitored, maintainable systems. Build model data foundations that ensure training, inference, monitoring, and analytics data are trustworthy and scalable. Architect ML platform patterns in Databricks that bring reliability, consistency, governance, performance, and cost discipline to ML and data workflows. Identify and scope opportunities for ML engineering across the business for high-impact. Develop reusable tools , libraries, standards, documentation, and production-readiness practices to enable data science and data engineering teams. Develop analytical and model-powered applications that turn data and ML outputs into usable business workflows for end users. Prepare the platform for future AI engineering , including LLM and agent-based systems, as the organization matures. Provide technical leadership and mentoring across engineering, architecture, and development including design and code reviews. Job Requirements: Technical Skills: Quantitative Foundation : B.S. degree in a quantitative field (e.g., Computer Science, Mathematics, Statistics, Economics, Operations Research, Engineering). Software Engineering Fundamentals: write clean, modular, testable, maintainable code and understand how to structure production-grade systems rather than one-off notebooks or scripts. Python and SQL Proficiency: strong in Python and SQL for building data pipelines, automation, model integrations, analytical workflows, and production services. Data Modeling and Pipeline Design: understand how to design reliable, well-structured data assets, including curated tables, feature datasets, batch pipelines, orchestration, data quality checks, and lineage. ML Lifecycle Fluency: understand the full model lifecycle: data collection, exploration, model development, validation, deployment, monitoring, retraining, and retirement. Production ML Patterns: understand core MLOps patterns such as model registries, feature/data versioning, reproducible environments, testing/validation, monitoring, and rollback. Cloud and Platform Engineering: You are comfortable working in cloud-based data and ML environments and understand the foundations of permissions, environments, jobs, services, storage, networking, and cost-aware architecture. Databricks Expertise : You're familiar and experienced with the core parts of Spark, Unity Catalog, Delta Lake, Databricks Workflows, MLflow, model registry patterns, job/cluster optimization, and governance. DevOps Practices: You use modern engineering practices such as Git, CI/CD, automated testing, code review, dependency management, environment management, and observability. Application Development : You can build applications, APIs, dashboards, or workflow tools that sit on top of data and model outputs. System Design: You can reason through tradeoffs across reliability, latency, scale, cost, governance, maintainability, and ease of use. Soft Skills: Curious : bring intellectual curiosity, an inquisitive nature, and a desire to deepen your knowledge and continue learning. Ownership : take responsibility to proactively advance projects, contribute to the organization, and develop the best solutions. Communication: explain technical concepts, risks, tradeoffs, and recommendations clearly to technical and non-technical audiences. Collaboration: work effectively cross-functionally with data scientists, data engineers, analysts, application engineers, product partners, and business stakeholders. Pragmatism : You know how to balance ideal architecture with business urgency, team maturity, operational constraints, and the need to ship. Preferred qualifications: A graduate degree (Masters or PhD) in a quantitative field Experience with dbt (Core) for modular data modeling, including testing, documentation, and dependency management Experience with AI engineer to use, build, and monitor agentic solutions The expected Total Compensation for this role is $140,000 - $185,000 + Annual Bonus. Individual compensation decisions are based on a variety of factors. Job Benefits Ski/Mountain Perks! Free passes for employees, employee discounted lift tickets for friends and family AND free ski lessons MORE employee discounts on lodging, food, gear, and mountain shuttles 401(k) Retirement Plan Employee Assistance Program Excellent training and professional development Full Time roles are eligible for the above, plus: Health Insurance; Medical Insurance, Dental Insurance, and Vision Insurance plans (for eligible seasonal employees after working 500 hours) Free ski passes for dependents Critical Illness and Accident plans Employees can work remotely from British Columbia, Washington D.C., and the 16 U.S. states in which we currently operate. This includes: California, Colorado, Indiana, Michigan, Minnesota, Missouri, New Hampshire, New York, Nevada, Ohio, Pennsylvania, Utah, Vermont, Washington State, Wisconsin, and Wyoming. Please note that the ability to work in person or off-site, and the particulars related to such work, are subject to change at any time; and, accordingly, the Company reserves the right to change its policies and/or require in-person/in-office work or off-site work at any time in its sole discretion. In completing this application, and when submitting related documentation, applicants may redact information that identifies their age, date of birth, and/or dates of attendance at or graduation from an educational institution. We follow all federal, state, and local laws including restrictions on child/minor labor. Minors hired into this position will not be asked or permitted to engage in any activities restricted to adult workers. Vail Resorts is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability, protected veteran status or any other status protected by applicable law. Requisition ID 517322 Reference Date: 09/05/2026 Job Code Function: Data Science
09/19/2026
Full time
Job Description Job Description Our mission is to create the Experience of a Lifetime for our employees, so they can, in turn, create the Experience of a Lifetime for our guests. We own and operate the most renowned destination resorts in the world as well as regional and local ski areas outside major cities, and connect them all through one unrivaled network. We are looking for ambitious leaders, innovators and creators to join our talented team. If you're ready to pursue your fullest potential, we want to get to know you! Candidates for year-round positions are reviewed on a rolling basis. Applications will be accepted up to 90 days after the posting date, or until the position is filled (whichever is first). Job Summary: We are looking for a curious, driven, innovative machine learning engineer who takes initiative to solve problems and create environments that accelerate the development, deployment, and usage of data science models and AI to drive greater organizational impact. The Data Science & Data Engineering team within the Enterprise Analytics organization builds data assets, predictive models, analytical applications, and platforms across the organization. Our team collaborates with business stakeholders, analysts, and technology teams to tackle high-impact use cases with state-of-the-art models and tools to grow the business, streamline costs, and improve guest experiences. Job Specifications: Starting Wage: $140,000 - $185,000 + Annual Bonus Employment Type: Year Round Shift Type: Full Time hours Minimum Age: At least 18 years of age Housing Availability: No Job Responsibilities: Productionize ML models developed by data science into reliable, monitored, maintainable systems. Build model data foundations that ensure training, inference, monitoring, and analytics data are trustworthy and scalable. Architect ML platform patterns in Databricks that bring reliability, consistency, governance, performance, and cost discipline to ML and data workflows. Identify and scope opportunities for ML engineering across the business for high-impact. Develop reusable tools , libraries, standards, documentation, and production-readiness practices to enable data science and data engineering teams. Develop analytical and model-powered applications that turn data and ML outputs into usable business workflows for end users. Prepare the platform for future AI engineering , including LLM and agent-based systems, as the organization matures. Provide technical leadership and mentoring across engineering, architecture, and development including design and code reviews. Job Requirements: Technical Skills: Quantitative Foundation : B.S. degree in a quantitative field (e.g., Computer Science, Mathematics, Statistics, Economics, Operations Research, Engineering). Software Engineering Fundamentals: write clean, modular, testable, maintainable code and understand how to structure production-grade systems rather than one-off notebooks or scripts. Python and SQL Proficiency: strong in Python and SQL for building data pipelines, automation, model integrations, analytical workflows, and production services. Data Modeling and Pipeline Design: understand how to design reliable, well-structured data assets, including curated tables, feature datasets, batch pipelines, orchestration, data quality checks, and lineage. ML Lifecycle Fluency: understand the full model lifecycle: data collection, exploration, model development, validation, deployment, monitoring, retraining, and retirement. Production ML Patterns: understand core MLOps patterns such as model registries, feature/data versioning, reproducible environments, testing/validation, monitoring, and rollback. Cloud and Platform Engineering: You are comfortable working in cloud-based data and ML environments and understand the foundations of permissions, environments, jobs, services, storage, networking, and cost-aware architecture. Databricks Expertise : You're familiar and experienced with the core parts of Spark, Unity Catalog, Delta Lake, Databricks Workflows, MLflow, model registry patterns, job/cluster optimization, and governance. DevOps Practices: You use modern engineering practices such as Git, CI/CD, automated testing, code review, dependency management, environment management, and observability. Application Development : You can build applications, APIs, dashboards, or workflow tools that sit on top of data and model outputs. System Design: You can reason through tradeoffs across reliability, latency, scale, cost, governance, maintainability, and ease of use. Soft Skills: Curious : bring intellectual curiosity, an inquisitive nature, and a desire to deepen your knowledge and continue learning. Ownership : take responsibility to proactively advance projects, contribute to the organization, and develop the best solutions. Communication: explain technical concepts, risks, tradeoffs, and recommendations clearly to technical and non-technical audiences. Collaboration: work effectively cross-functionally with data scientists, data engineers, analysts, application engineers, product partners, and business stakeholders. Pragmatism : You know how to balance ideal architecture with business urgency, team maturity, operational constraints, and the need to ship. Preferred qualifications: A graduate degree (Masters or PhD) in a quantitative field Experience with dbt (Core) for modular data modeling, including testing, documentation, and dependency management Experience with AI engineer to use, build, and monitor agentic solutions The expected Total Compensation for this role is $140,000 - $185,000 + Annual Bonus. Individual compensation decisions are based on a variety of factors. Job Benefits Ski/Mountain Perks! Free passes for employees, employee discounted lift tickets for friends and family AND free ski lessons MORE employee discounts on lodging, food, gear, and mountain shuttles 401(k) Retirement Plan Employee Assistance Program Excellent training and professional development Full Time roles are eligible for the above, plus: Health Insurance; Medical Insurance, Dental Insurance, and Vision Insurance plans (for eligible seasonal employees after working 500 hours) Free ski passes for dependents Critical Illness and Accident plans Employees can work remotely from British Columbia, Washington D.C., and the 16 U.S. states in which we currently operate. This includes: California, Colorado, Indiana, Michigan, Minnesota, Missouri, New Hampshire, New York, Nevada, Ohio, Pennsylvania, Utah, Vermont, Washington State, Wisconsin, and Wyoming. Please note that the ability to work in person or off-site, and the particulars related to such work, are subject to change at any time; and, accordingly, the Company reserves the right to change its policies and/or require in-person/in-office work or off-site work at any time in its sole discretion. In completing this application, and when submitting related documentation, applicants may redact information that identifies their age, date of birth, and/or dates of attendance at or graduation from an educational institution. We follow all federal, state, and local laws including restrictions on child/minor labor. Minors hired into this position will not be asked or permitted to engage in any activities restricted to adult workers. Vail Resorts is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability, protected veteran status or any other status protected by applicable law. Requisition ID 517322 Reference Date: 09/05/2026 Job Code Function: Data Science
Job Description Job Description About the Company : Sungrow North America is a leading provider of renewable energy solutions, specializing in the development and manufacturing of photovoltaic inverters and energy storage systems. The company offers a comprehensive range of products and services designed to optimize the performance and efficiency of solar power installations. Sungrow North America aims to provide sustainable and reliable energy solutions to meet the growing demand for clean power and is known for its commitment to innovation, high-quality standards, and exceptional customer service. Security Engineer - Network & Identity: The Security Engineer (Network & Identity) is a hands-on engineering role within the IT team responsible for designing, implementing, securing, and automating Sungrow USA's network security, PKI and certificate management, and identity & access infrastructure across on-premises, cloud, and SaaS environments. This role serves as the technical owner for network security architecture, cryptographic services, certificate lifecycle management, authentication, and access controls. The position focuses on Zero Trust security, network segmentation, certificate-based authentication, and identity protection to reduce organizational risk and enable secure business operations and platform ownership rather than SOC operations, threat monitoring, or incident response. Essential Duties and Responsibilities: Network Security Design, implement, and maintain secure enterprise network architectures across corporate offices, data centers, cloud platforms, and remote workforce environments. Architect network segmentation, Zero Trust access controls, and secure connectivity standards using Fortinet and Zscaler security solutions. Develop and maintain Zero Trust architectures across network, identity, endpoint, application, and cloud environments. Design and administer secure remote access using Zscaler Private Access, VPN technologies, and identity-aware access controls. Manage firewall policies, network security controls, routing security, DNS security, and hybrid-cloud connectivity. Design and support Network Access Control architectures using IEEE 802.1X, RADIUS, and certificate-based authentication. Assess network security posture, develop remediation plans, and drive continuous security improvements. Cryptography, PKI & Certificate Management Own the enterprise PKI, cryptography, and certificate lifecycle management architecture, standards, and governance program. Design and manage certificate-based authentication and machine identity solutions for users, devices, servers, applications, cloud workloads, and network infrastructure across Azure, AWS, and hybrid environments. Implement and maintain certificate lifecycle automation using Microsoft Cloud PKI, Keyfactor, CyberArk Certificate Manager, EJBCA, DigiCert, AppViewX, or comparable platforms. Manage certificate issuance, enrollment, discovery, deployment, monitoring, renewal, revocation, auditing, and compliance across the enterprise. Design and support cryptographic services and certificate-based security controls, including TLS/mTLS, code signing, PKI trust hierarchies, certificate-based authentication, SCEP, PKCS, and machine identities. Establish PKI and cryptographic standards, key management practices, and security controls to support Zero Trust, regulatory compliance, and enterprise security requirements. Troubleshoot and resolve complex certificate, cryptographic, trust chain, authentication, and secure communications issues across enterprise systems and applications. Identity & Access Define authentication and authorization standards for workforce, partner, application, service, and machine identities. Design, implement, and maintain Microsoft Entra ID architecture, tenant governance, and identity security controls. Develop, test, and enforce Conditional Access policies and Zero Trust access controls. Implement and maintain MFA, passwordless authentication, phishing-resistant authentication, and Microsoft Entra ID Protection capabilities. Design and support enterprise SSO and federation integrations using SAML, OAuth 2.0, OpenID Connect, and SCIM. Implement least-privilege and risk-based access models across enterprise platforms. Administer RBAC, administrative separation, Microsoft Entra Privileged Identity Management, and least-privilege access controls. Govern application registrations, service principals, enterprise applications, API permissions, and managed identities. Support B2B collaboration, guest-user governance, external workforce access, and third-party identity integrations. Design and implement security controls across Microsoft Azure and AWS environments. Apply least privilege, RBAC, encryption, secrets management, and secure configuration standards to on-prem and cloud resources. Automate identity provisioning and deprovisioning, access governance, certificate management, configuration validation, and security operations. Create reusable secure-by-default templates and reduce manual administration through automation and orchestration Conduct access reviews, entitlement certifications, and identity governance activities. Education or Desired License and Certificates: Bachelor's degree in Computer Science, Information Technology, Cybersecurity, Engineering, or a related field, or equivalent professional experience. Microsoft Certified: Identity and Access Administrator Associate (SC-300) preferred. Microsoft Certified: Azure Security Engineer Associate (AZ-500) preferred. CCNA, Fortinet, Zscaler, AWS Security, CISSP, CISM, Terraform, or relevant PKI certification preferred Preferred Experience & Qualifications: 5+ years of experience in security engineering, identity & access management (IAM), network security, cloud security, or a related enterprise IT discipline. Hands-on experience with Microsoft Entra ID, including Conditional Access, MFA, SSO, Identity Protection, PIM, RBAC, identity governance, and modern authentication protocols (SAML, OAuth, OpenID Connect, SCIM). Experience designing, implementing, and securing enterprise identity, privileged access, and machine identity solutions across hybrid and multi-cloud environments. Hands-on experience with Fortinet, Zscaler (ZIA/ZPA), Zero Trust architectures, least-privilege access models, and network security controls. Experience designing and operating enterprise PKI, certificate lifecycle management, certificate-based authentication, and machine identity platforms such as Keyfactor, DigiCert, EJBCA, AppViewX, CyberArk Certificate Manager, or similar solutions. Experience securing Azure and AWS environments, including identity, networking, encryption, secrets management, logging, and security monitoring. Experience with PAM and IGA platforms such as CyberArk, Delinea, BeyondTrust, SailPoint, Saviynt, or similar technologies. Experience integrating identity, network, cloud, and security telemetry with SIEM and security operations platforms. Strong automation and Infrastructure as Code skills using PowerShell, Python, Microsoft Graph API, REST APIs, Terraform, or similar technologies. Strong troubleshooting skills across authentication, federation, certificates, PKI, network security, cloud access, application integrations, and enterprise identity services. Knowledge of cybersecurity and compliance frameworks including SOC 2, ISO/IEC 27001, NIST CSF, NIST 800-63, CIS Controls, Zero Trust, and NERC CIP. Competencies: Mandarin fluency preferred but not required. Strong analytical, troubleshooting, and problem-solving skills. Ability to work independently and collaboratively in a fast-paced environment. Excellent communication, stakeholder management, and technical documentation skills. Strong organization, attention to detail, initiative, and ownership. Ability to balance security, reliability, usability, scalability, and business requirements. Proactive approach to automation, standardization, and continuous improvement. Travel 5%-20% Work Location and Status: Full time, Hybrid at any Sungrow USA office in Phoenix, Costa Mesa, or Houston No visa sponsorship Compensation: Compensation commensurate with experience Competitive salary and annual bonus eligibility Comprehensive benefits package including health, dental, vision, and retirement plans Strong personal and company growth opportunities Sungrow is an equal opportunity employer. Due to strong interest in this position, Sungrow will only reach out to those candidates who best meet the requirements. Thank you for your interest in Sungrow.
09/19/2026
Full time
Job Description Job Description About the Company : Sungrow North America is a leading provider of renewable energy solutions, specializing in the development and manufacturing of photovoltaic inverters and energy storage systems. The company offers a comprehensive range of products and services designed to optimize the performance and efficiency of solar power installations. Sungrow North America aims to provide sustainable and reliable energy solutions to meet the growing demand for clean power and is known for its commitment to innovation, high-quality standards, and exceptional customer service. Security Engineer - Network & Identity: The Security Engineer (Network & Identity) is a hands-on engineering role within the IT team responsible for designing, implementing, securing, and automating Sungrow USA's network security, PKI and certificate management, and identity & access infrastructure across on-premises, cloud, and SaaS environments. This role serves as the technical owner for network security architecture, cryptographic services, certificate lifecycle management, authentication, and access controls. The position focuses on Zero Trust security, network segmentation, certificate-based authentication, and identity protection to reduce organizational risk and enable secure business operations and platform ownership rather than SOC operations, threat monitoring, or incident response. Essential Duties and Responsibilities: Network Security Design, implement, and maintain secure enterprise network architectures across corporate offices, data centers, cloud platforms, and remote workforce environments. Architect network segmentation, Zero Trust access controls, and secure connectivity standards using Fortinet and Zscaler security solutions. Develop and maintain Zero Trust architectures across network, identity, endpoint, application, and cloud environments. Design and administer secure remote access using Zscaler Private Access, VPN technologies, and identity-aware access controls. Manage firewall policies, network security controls, routing security, DNS security, and hybrid-cloud connectivity. Design and support Network Access Control architectures using IEEE 802.1X, RADIUS, and certificate-based authentication. Assess network security posture, develop remediation plans, and drive continuous security improvements. Cryptography, PKI & Certificate Management Own the enterprise PKI, cryptography, and certificate lifecycle management architecture, standards, and governance program. Design and manage certificate-based authentication and machine identity solutions for users, devices, servers, applications, cloud workloads, and network infrastructure across Azure, AWS, and hybrid environments. Implement and maintain certificate lifecycle automation using Microsoft Cloud PKI, Keyfactor, CyberArk Certificate Manager, EJBCA, DigiCert, AppViewX, or comparable platforms. Manage certificate issuance, enrollment, discovery, deployment, monitoring, renewal, revocation, auditing, and compliance across the enterprise. Design and support cryptographic services and certificate-based security controls, including TLS/mTLS, code signing, PKI trust hierarchies, certificate-based authentication, SCEP, PKCS, and machine identities. Establish PKI and cryptographic standards, key management practices, and security controls to support Zero Trust, regulatory compliance, and enterprise security requirements. Troubleshoot and resolve complex certificate, cryptographic, trust chain, authentication, and secure communications issues across enterprise systems and applications. Identity & Access Define authentication and authorization standards for workforce, partner, application, service, and machine identities. Design, implement, and maintain Microsoft Entra ID architecture, tenant governance, and identity security controls. Develop, test, and enforce Conditional Access policies and Zero Trust access controls. Implement and maintain MFA, passwordless authentication, phishing-resistant authentication, and Microsoft Entra ID Protection capabilities. Design and support enterprise SSO and federation integrations using SAML, OAuth 2.0, OpenID Connect, and SCIM. Implement least-privilege and risk-based access models across enterprise platforms. Administer RBAC, administrative separation, Microsoft Entra Privileged Identity Management, and least-privilege access controls. Govern application registrations, service principals, enterprise applications, API permissions, and managed identities. Support B2B collaboration, guest-user governance, external workforce access, and third-party identity integrations. Design and implement security controls across Microsoft Azure and AWS environments. Apply least privilege, RBAC, encryption, secrets management, and secure configuration standards to on-prem and cloud resources. Automate identity provisioning and deprovisioning, access governance, certificate management, configuration validation, and security operations. Create reusable secure-by-default templates and reduce manual administration through automation and orchestration Conduct access reviews, entitlement certifications, and identity governance activities. Education or Desired License and Certificates: Bachelor's degree in Computer Science, Information Technology, Cybersecurity, Engineering, or a related field, or equivalent professional experience. Microsoft Certified: Identity and Access Administrator Associate (SC-300) preferred. Microsoft Certified: Azure Security Engineer Associate (AZ-500) preferred. CCNA, Fortinet, Zscaler, AWS Security, CISSP, CISM, Terraform, or relevant PKI certification preferred Preferred Experience & Qualifications: 5+ years of experience in security engineering, identity & access management (IAM), network security, cloud security, or a related enterprise IT discipline. Hands-on experience with Microsoft Entra ID, including Conditional Access, MFA, SSO, Identity Protection, PIM, RBAC, identity governance, and modern authentication protocols (SAML, OAuth, OpenID Connect, SCIM). Experience designing, implementing, and securing enterprise identity, privileged access, and machine identity solutions across hybrid and multi-cloud environments. Hands-on experience with Fortinet, Zscaler (ZIA/ZPA), Zero Trust architectures, least-privilege access models, and network security controls. Experience designing and operating enterprise PKI, certificate lifecycle management, certificate-based authentication, and machine identity platforms such as Keyfactor, DigiCert, EJBCA, AppViewX, CyberArk Certificate Manager, or similar solutions. Experience securing Azure and AWS environments, including identity, networking, encryption, secrets management, logging, and security monitoring. Experience with PAM and IGA platforms such as CyberArk, Delinea, BeyondTrust, SailPoint, Saviynt, or similar technologies. Experience integrating identity, network, cloud, and security telemetry with SIEM and security operations platforms. Strong automation and Infrastructure as Code skills using PowerShell, Python, Microsoft Graph API, REST APIs, Terraform, or similar technologies. Strong troubleshooting skills across authentication, federation, certificates, PKI, network security, cloud access, application integrations, and enterprise identity services. Knowledge of cybersecurity and compliance frameworks including SOC 2, ISO/IEC 27001, NIST CSF, NIST 800-63, CIS Controls, Zero Trust, and NERC CIP. Competencies: Mandarin fluency preferred but not required. Strong analytical, troubleshooting, and problem-solving skills. Ability to work independently and collaboratively in a fast-paced environment. Excellent communication, stakeholder management, and technical documentation skills. Strong organization, attention to detail, initiative, and ownership. Ability to balance security, reliability, usability, scalability, and business requirements. Proactive approach to automation, standardization, and continuous improvement. Travel 5%-20% Work Location and Status: Full time, Hybrid at any Sungrow USA office in Phoenix, Costa Mesa, or Houston No visa sponsorship Compensation: Compensation commensurate with experience Competitive salary and annual bonus eligibility Comprehensive benefits package including health, dental, vision, and retirement plans Strong personal and company growth opportunities Sungrow is an equal opportunity employer. Due to strong interest in this position, Sungrow will only reach out to those candidates who best meet the requirements. Thank you for your interest in Sungrow.
AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada . click apply for full job details
09/19/2026
Full time
AI Engineer 5 (FM Hosting, LLM Inference) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada . click apply for full job details
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 5 (SDK's: Gen AI Evaluation and MCP) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities . click apply for full job details
09/19/2026
Full time
AI Engineer 5 (SDK's: Gen AI Evaluation and MCP) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities . click apply for full job details
AI Engineer 5 (Gen AI Platform Services - Agentic AI) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities . click apply for full job details
09/19/2026
Full time
AI Engineer 5 (Gen AI Platform Services - Agentic AI) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities . click apply for full job details
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 5 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada . click apply for full job details
09/19/2026
Full time
AI Engineer 5 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada . click apply for full job details
AI Engineer 4 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $197,300 - $225,100 for AI Engineer 4 Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $197,300 - $225,100 for AI Engineer 4 Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 5 (AI Foundations) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $229,900 - $262,400 for AI Engineer 5 Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada . click apply for full job details
09/19/2026
Full time
AI Engineer 5 (AI Foundations) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $229,900 - $262,400 for AI Engineer 5 Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada . click apply for full job details
AI Engineer 5 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities . click apply for full job details
09/19/2026
Full time
AI Engineer 5 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Design, implement and optimize multi-model orchestration pipelines - integrating LLMs, vector search, and domain-specific models into unified systems Establish and lead cost-performance governance reviews across AI systems, tracking GPU utilization, model throughput, and inference cost efficiency Lead team design councils or design review boards to ensure technical consistency and compliance with AI engineering standards Mentor Principal and Manager-level AI engineers, fostering cross-domain learning and elevating organizational technical maturity The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting complex AI systems Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Experience architecting and integrating heterogeneous AI systems - including rule-based, retrieval-augmented, and generative components - into unified production pipelines Experience defining and enforcing standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes Demonstrated ability to balance model performance and operational cost through dynamic inference strategies and model compression Experience right-sizing models, instance counts, and hardware types given requirements (e.g., context length, token inputs, token outputs) Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for AI Engineer 5 McLean, VA: $229,900 - $262,400 for AI Engineer 5 New York, NY: $250,800 - $286,200 for AI Engineer 5 San Francisco, CA: $250,800 - $286,200 for AI Engineer 5 San Jose, CA: $250,800 - $286,200 for AI Engineer 5 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities . click apply for full job details