Job Description Job Description Material Bank is the world's largest material marketplace for the architecture and design industry. Operating in 37 countries, our platform has become the standard for design professionals around the globe. Every day, Material Bank connects thousands of designers with tens of thousands of materials from leading brands. Material Bank is the fastest and most powerful way for design professionals to search, sample, and specify materials. About the role As an Applied AI Engineer, you will drive the design, build, and deployment of next-generation AI-powered experiences across Material Bank's platform. You will work as part of the team responsible for taking ideas from concept to production - building intelligent systems and user experiences that blend cutting-edge AI capabilities with the high standards of quality, aesthetics, and usability expected in the architecture and community. This is a senior-level individual contributor role focused on applied AI product development. You will work across the stack to architect and deploy scalable AI systems that enhance how users discover, understand, and engage with products, materials, and creative content. Your work will span areas such as multimodal search and understanding, AI-assisted content generation, intelligent workflows, personalization, creative tooling, and agentic systems. We are looking for someone who not only understands modern AI systems technically, but also has strong product instincts, visual sensibility, and genuine passion for building AI experiences that feel thoughtful, polished, and useful. AI will play a foundational role in the future of Material Bank's platform, and you will help define and build that future. What you'll do Design, build, and deploy end-to-end AI-powered product experiences from concept through production. Architect and implement scalable AI systems leveraging LLMs, embeddings, multimodal models, retrieval systems, agent frameworks, and modern data infrastructure. Build production-grade multi-agent workflows and orchestration systems using frameworks such as LangGraph, LangChain, Mastra, and custom tooling. Develop and optimize Retrieval-Augmented Generation (RAG) systems, including embeddings, vector search, retrieval pipelines, chunking strategies, and relevance tuning. Build multimodal AI workflows that analyze and reason over images, creative assets, and visual datasets using modern multimodal LLMs, embedding models, and specialized tooling such as SAM2/SAM3. Create AI-assisted experiences for search, discovery, content generation, personalization, and creative workflows across Material Bank's platform. Evaluate, refine, and improve AI-generated outputs for quality, tone, accuracy, and creative alignment through testing, iteration, and human-in-the-loop evaluation strategies. Partner closely with Product, Design, Engineering, Data, and Executive Leadership to identify high-impact opportunities and translate ambiguous ideas into production-ready AI capabilities. Make architectural decisions that balance speed, scalability, latency, cost, accuracy, and long-term maintainability. Continuously evaluate emerging AI technologies, models, frameworks, and workflows to identify opportunities that create meaningful business and user value. What you'll bring 8+ years of experience building and shipping production software, including significant full-stack engineering experience. Demonstrated success designing and deploying production-grade AI/ML systems and AI-powered product experiences. Deep hands-on experience with LLMs, embeddings, multimodal AI systems, RAG architectures, and multi-agent frameworks such as LangGraph, LangChain, Mastra, or equivalent custom tooling. Strong engineering fundamentals across backend systems, APIs, data pipelines, cloud infrastructure, and modern JavaScript/TypeScript and Python ecosystems. Experience working with multimodal models, visual analysis systems, and image-based AI workflows at scale, including familiarity with modern image-generation tooling and services. Strong systems thinking with the ability to balance trade-offs across latency, cost, scalability, accuracy, reliability, and user experience. Proven ability to independently take ambiguous problems from idea to shipped product with minimal oversight. Strong product instincts, visual sensibility, and a high bar for quality, usability, and craftsmanship in AI-generated experiences. Genuine interest in creative industries such as architecture, design, fashion, media, photography, or art, with an appreciation for aesthetics and taste. Open-source contributions, side projects, or publicly demonstrable AI work that reflects curiosity, experimentation, and passion for applied AI are strongly preferred. Strong communication and collaboration skills, with the ability to work effectively across both technical and non-technical teams. What you'll get from us: Our people : We are a growth-driven team that values efficiency, builds smart automation, operates in small empowered teams, and moves quickly from idea to execution. Relaxation and Celebrations : Flexible PTO, Sick Days, Paid National Holidays, and even more (ask us about this when we connect). Health Benefits : We contribute to your medical, dental, vision and short-term/long-term disability plans and have a strong employee assistance program. Plan for your Retirement : 401(k) eligible after your first 90 day's employed! Giving Back : We sponsor multiple events throughout the year to help out our communities. Growth : We'll help you take your career to the next level. We want you to be creative and take initiative which will allow you to grow and create within the company. Most importantly, be the best at what matters! Flexible Work Schedules : With business units and employees across the globe, Material Technologies has embraced a hybrid working model allowing department leaders to decide on the best approach for their respective teams, whether that be remote, in person, or a little of both. Material Bank is proud to be an equal opportunity employer. We value diversity, and all applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, age, national origin, veteran or disability status or other status protected under any applicable federal, state or local law.
09/20/2026
Full time
Job Description Job Description Material Bank is the world's largest material marketplace for the architecture and design industry. Operating in 37 countries, our platform has become the standard for design professionals around the globe. Every day, Material Bank connects thousands of designers with tens of thousands of materials from leading brands. Material Bank is the fastest and most powerful way for design professionals to search, sample, and specify materials. About the role As an Applied AI Engineer, you will drive the design, build, and deployment of next-generation AI-powered experiences across Material Bank's platform. You will work as part of the team responsible for taking ideas from concept to production - building intelligent systems and user experiences that blend cutting-edge AI capabilities with the high standards of quality, aesthetics, and usability expected in the architecture and community. This is a senior-level individual contributor role focused on applied AI product development. You will work across the stack to architect and deploy scalable AI systems that enhance how users discover, understand, and engage with products, materials, and creative content. Your work will span areas such as multimodal search and understanding, AI-assisted content generation, intelligent workflows, personalization, creative tooling, and agentic systems. We are looking for someone who not only understands modern AI systems technically, but also has strong product instincts, visual sensibility, and genuine passion for building AI experiences that feel thoughtful, polished, and useful. AI will play a foundational role in the future of Material Bank's platform, and you will help define and build that future. What you'll do Design, build, and deploy end-to-end AI-powered product experiences from concept through production. Architect and implement scalable AI systems leveraging LLMs, embeddings, multimodal models, retrieval systems, agent frameworks, and modern data infrastructure. Build production-grade multi-agent workflows and orchestration systems using frameworks such as LangGraph, LangChain, Mastra, and custom tooling. Develop and optimize Retrieval-Augmented Generation (RAG) systems, including embeddings, vector search, retrieval pipelines, chunking strategies, and relevance tuning. Build multimodal AI workflows that analyze and reason over images, creative assets, and visual datasets using modern multimodal LLMs, embedding models, and specialized tooling such as SAM2/SAM3. Create AI-assisted experiences for search, discovery, content generation, personalization, and creative workflows across Material Bank's platform. Evaluate, refine, and improve AI-generated outputs for quality, tone, accuracy, and creative alignment through testing, iteration, and human-in-the-loop evaluation strategies. Partner closely with Product, Design, Engineering, Data, and Executive Leadership to identify high-impact opportunities and translate ambiguous ideas into production-ready AI capabilities. Make architectural decisions that balance speed, scalability, latency, cost, accuracy, and long-term maintainability. Continuously evaluate emerging AI technologies, models, frameworks, and workflows to identify opportunities that create meaningful business and user value. What you'll bring 8+ years of experience building and shipping production software, including significant full-stack engineering experience. Demonstrated success designing and deploying production-grade AI/ML systems and AI-powered product experiences. Deep hands-on experience with LLMs, embeddings, multimodal AI systems, RAG architectures, and multi-agent frameworks such as LangGraph, LangChain, Mastra, or equivalent custom tooling. Strong engineering fundamentals across backend systems, APIs, data pipelines, cloud infrastructure, and modern JavaScript/TypeScript and Python ecosystems. Experience working with multimodal models, visual analysis systems, and image-based AI workflows at scale, including familiarity with modern image-generation tooling and services. Strong systems thinking with the ability to balance trade-offs across latency, cost, scalability, accuracy, reliability, and user experience. Proven ability to independently take ambiguous problems from idea to shipped product with minimal oversight. Strong product instincts, visual sensibility, and a high bar for quality, usability, and craftsmanship in AI-generated experiences. Genuine interest in creative industries such as architecture, design, fashion, media, photography, or art, with an appreciation for aesthetics and taste. Open-source contributions, side projects, or publicly demonstrable AI work that reflects curiosity, experimentation, and passion for applied AI are strongly preferred. Strong communication and collaboration skills, with the ability to work effectively across both technical and non-technical teams. What you'll get from us: Our people : We are a growth-driven team that values efficiency, builds smart automation, operates in small empowered teams, and moves quickly from idea to execution. Relaxation and Celebrations : Flexible PTO, Sick Days, Paid National Holidays, and even more (ask us about this when we connect). Health Benefits : We contribute to your medical, dental, vision and short-term/long-term disability plans and have a strong employee assistance program. Plan for your Retirement : 401(k) eligible after your first 90 day's employed! Giving Back : We sponsor multiple events throughout the year to help out our communities. Growth : We'll help you take your career to the next level. We want you to be creative and take initiative which will allow you to grow and create within the company. Most importantly, be the best at what matters! Flexible Work Schedules : With business units and employees across the globe, Material Technologies has embraced a hybrid working model allowing department leaders to decide on the best approach for their respective teams, whether that be remote, in person, or a little of both. Material Bank is proud to be an equal opportunity employer. We value diversity, and all applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, age, national origin, veteran or disability status or other status protected under any applicable federal, state or local law.
Job Description Job Description Material Bank is the world's largest material marketplace for the architecture and design industry. Operating in 37 countries, our platform has become the standard for design professionals around the globe. Every day, Material Bank connects thousands of designers with tens of thousands of materials from leading brands. Material Bank is the fastest and most powerful way for design professionals to search, sample, and specify materials. About the role As an Applied AI Engineer, you will drive the design, build, and deployment of next-generation AI-powered experiences across Material Bank's platform. You will work as part of the team responsible for taking ideas from concept to production - building intelligent systems and user experiences that blend cutting-edge AI capabilities with the high standards of quality, aesthetics, and usability expected in the architecture and community. This is a senior-level individual contributor role focused on applied AI product development. You will work across the stack to architect and deploy scalable AI systems that enhance how users discover, understand, and engage with products, materials, and creative content. Your work will span areas such as multimodal search and understanding, AI-assisted content generation, intelligent workflows, personalization, creative tooling, and agentic systems. We are looking for someone who not only understands modern AI systems technically, but also has strong product instincts, visual sensibility, and genuine passion for building AI experiences that feel thoughtful, polished, and useful. AI will play a foundational role in the future of Material Bank's platform, and you will help define and build that future. What you'll do Design, build, and deploy end-to-end AI-powered product experiences from concept through production. Architect and implement scalable AI systems leveraging LLMs, embeddings, multimodal models, retrieval systems, agent frameworks, and modern data infrastructure. Build production-grade multi-agent workflows and orchestration systems using frameworks such as LangGraph, LangChain, Mastra, and custom tooling. Develop and optimize Retrieval-Augmented Generation (RAG) systems, including embeddings, vector search, retrieval pipelines, chunking strategies, and relevance tuning. Build multimodal AI workflows that analyze and reason over images, creative assets, and visual datasets using modern multimodal LLMs, embedding models, and specialized tooling such as SAM2/SAM3. Create AI-assisted experiences for search, discovery, content generation, personalization, and creative workflows across Material Bank's platform. Evaluate, refine, and improve AI-generated outputs for quality, tone, accuracy, and creative alignment through testing, iteration, and human-in-the-loop evaluation strategies. Partner closely with Product, Design, Engineering, Data, and Executive Leadership to identify high-impact opportunities and translate ambiguous ideas into production-ready AI capabilities. Make architectural decisions that balance speed, scalability, latency, cost, accuracy, and long-term maintainability. Continuously evaluate emerging AI technologies, models, frameworks, and workflows to identify opportunities that create meaningful business and user value. What you'll bring 8+ years of experience building and shipping production software, including significant full-stack engineering experience. Demonstrated success designing and deploying production-grade AI/ML systems and AI-powered product experiences. Deep hands-on experience with LLMs, embeddings, multimodal AI systems, RAG architectures, and multi-agent frameworks such as LangGraph, LangChain, Mastra, or equivalent custom tooling. Strong engineering fundamentals across backend systems, APIs, data pipelines, cloud infrastructure, and modern JavaScript/TypeScript and Python ecosystems. Experience working with multimodal models, visual analysis systems, and image-based AI workflows at scale, including familiarity with modern image-generation tooling and services. Strong systems thinking with the ability to balance trade-offs across latency, cost, scalability, accuracy, reliability, and user experience. Proven ability to independently take ambiguous problems from idea to shipped product with minimal oversight. Strong product instincts, visual sensibility, and a high bar for quality, usability, and craftsmanship in AI-generated experiences. Genuine interest in creative industries such as architecture, design, fashion, media, photography, or art, with an appreciation for aesthetics and taste. Open-source contributions, side projects, or publicly demonstrable AI work that reflects curiosity, experimentation, and passion for applied AI are strongly preferred. Strong communication and collaboration skills, with the ability to work effectively across both technical and non-technical teams. What you'll get from us: Our people : We are a growth-driven team that values efficiency, builds smart automation, operates in small empowered teams, and moves quickly from idea to execution. Relaxation and Celebrations : Flexible PTO, Sick Days, Paid National Holidays, and even more (ask us about this when we connect). Health Benefits : We contribute to your medical, dental, vision and short-term/long-term disability plans and have a strong employee assistance program. Plan for your Retirement : 401(k) eligible after your first 90 day's employed! Giving Back : We sponsor multiple events throughout the year to help out our communities. Growth : We'll help you take your career to the next level. We want you to be creative and take initiative which will allow you to grow and create within the company. Most importantly, be the best at what matters! Flexible Work Schedules : With business units and employees across the globe, Material Technologies has embraced a hybrid working model allowing department leaders to decide on the best approach for their respective teams, whether that be remote, in person, or a little of both. Material Bank is proud to be an equal opportunity employer. We value diversity, and all applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, age, national origin, veteran or disability status or other status protected under any applicable federal, state or local law.
09/20/2026
Full time
Job Description Job Description Material Bank is the world's largest material marketplace for the architecture and design industry. Operating in 37 countries, our platform has become the standard for design professionals around the globe. Every day, Material Bank connects thousands of designers with tens of thousands of materials from leading brands. Material Bank is the fastest and most powerful way for design professionals to search, sample, and specify materials. About the role As an Applied AI Engineer, you will drive the design, build, and deployment of next-generation AI-powered experiences across Material Bank's platform. You will work as part of the team responsible for taking ideas from concept to production - building intelligent systems and user experiences that blend cutting-edge AI capabilities with the high standards of quality, aesthetics, and usability expected in the architecture and community. This is a senior-level individual contributor role focused on applied AI product development. You will work across the stack to architect and deploy scalable AI systems that enhance how users discover, understand, and engage with products, materials, and creative content. Your work will span areas such as multimodal search and understanding, AI-assisted content generation, intelligent workflows, personalization, creative tooling, and agentic systems. We are looking for someone who not only understands modern AI systems technically, but also has strong product instincts, visual sensibility, and genuine passion for building AI experiences that feel thoughtful, polished, and useful. AI will play a foundational role in the future of Material Bank's platform, and you will help define and build that future. What you'll do Design, build, and deploy end-to-end AI-powered product experiences from concept through production. Architect and implement scalable AI systems leveraging LLMs, embeddings, multimodal models, retrieval systems, agent frameworks, and modern data infrastructure. Build production-grade multi-agent workflows and orchestration systems using frameworks such as LangGraph, LangChain, Mastra, and custom tooling. Develop and optimize Retrieval-Augmented Generation (RAG) systems, including embeddings, vector search, retrieval pipelines, chunking strategies, and relevance tuning. Build multimodal AI workflows that analyze and reason over images, creative assets, and visual datasets using modern multimodal LLMs, embedding models, and specialized tooling such as SAM2/SAM3. Create AI-assisted experiences for search, discovery, content generation, personalization, and creative workflows across Material Bank's platform. Evaluate, refine, and improve AI-generated outputs for quality, tone, accuracy, and creative alignment through testing, iteration, and human-in-the-loop evaluation strategies. Partner closely with Product, Design, Engineering, Data, and Executive Leadership to identify high-impact opportunities and translate ambiguous ideas into production-ready AI capabilities. Make architectural decisions that balance speed, scalability, latency, cost, accuracy, and long-term maintainability. Continuously evaluate emerging AI technologies, models, frameworks, and workflows to identify opportunities that create meaningful business and user value. What you'll bring 8+ years of experience building and shipping production software, including significant full-stack engineering experience. Demonstrated success designing and deploying production-grade AI/ML systems and AI-powered product experiences. Deep hands-on experience with LLMs, embeddings, multimodal AI systems, RAG architectures, and multi-agent frameworks such as LangGraph, LangChain, Mastra, or equivalent custom tooling. Strong engineering fundamentals across backend systems, APIs, data pipelines, cloud infrastructure, and modern JavaScript/TypeScript and Python ecosystems. Experience working with multimodal models, visual analysis systems, and image-based AI workflows at scale, including familiarity with modern image-generation tooling and services. Strong systems thinking with the ability to balance trade-offs across latency, cost, scalability, accuracy, reliability, and user experience. Proven ability to independently take ambiguous problems from idea to shipped product with minimal oversight. Strong product instincts, visual sensibility, and a high bar for quality, usability, and craftsmanship in AI-generated experiences. Genuine interest in creative industries such as architecture, design, fashion, media, photography, or art, with an appreciation for aesthetics and taste. Open-source contributions, side projects, or publicly demonstrable AI work that reflects curiosity, experimentation, and passion for applied AI are strongly preferred. Strong communication and collaboration skills, with the ability to work effectively across both technical and non-technical teams. What you'll get from us: Our people : We are a growth-driven team that values efficiency, builds smart automation, operates in small empowered teams, and moves quickly from idea to execution. Relaxation and Celebrations : Flexible PTO, Sick Days, Paid National Holidays, and even more (ask us about this when we connect). Health Benefits : We contribute to your medical, dental, vision and short-term/long-term disability plans and have a strong employee assistance program. Plan for your Retirement : 401(k) eligible after your first 90 day's employed! Giving Back : We sponsor multiple events throughout the year to help out our communities. Growth : We'll help you take your career to the next level. We want you to be creative and take initiative which will allow you to grow and create within the company. Most importantly, be the best at what matters! Flexible Work Schedules : With business units and employees across the globe, Material Technologies has embraced a hybrid working model allowing department leaders to decide on the best approach for their respective teams, whether that be remote, in person, or a little of both. Material Bank is proud to be an equal opportunity employer. We value diversity, and all applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, age, national origin, veteran or disability status or other status protected under any applicable federal, state or local law.
Job Description Job Description Genesis10 is currently seeking a Innovation Principal Engineer for a Contract-to-Hire opportunity. This position can work hybrid or can work remotely in the footprint locations of Columbus, OH; Minnetonka, MN; Atlanta, GA: Dallas, TX; Detroit, MI; Chicago, IL; Charlotte, NC; Birmingham, AL; Tupelo, MS; Pittsburgh, PA; Cleveland, OH; Akron, OH and Cincinnati, OH. Compensation: $100.00 - $110.00 per hour, W2, based on qualifications Position Overview: This position will be responsible for driving this organizations most complex innovation initiatives by designing and building end-to-end, production-grade prototypes and reusable technology assets, including AI-enabled and agentic applications within Google Cloud, while setting high engineering standards for solution architecture, application design, security, governance, and delivery acceleration across the Technology Innovation portfolio. This role requires strong hands-on full-stack engineering skills, the mindset of a startup innovator in a large organization, deep experience building AI-integrated applications end-to-end, developing APIs and user experiences, and the ability to operate independently with limited supervision in a fast-paced innovation environment. The ideal candidate will bring demonstrated production experience with Vertex AI, Gemini, agent frameworks, AI gateway/model-proxy patterns, and enterprise-grade deployment of AI workloads. Key Responsibilities: Lead the design and delivery of high-impact innovation solutions aligned with the organization's business goals and technology trends in financial services Architect and build end-to-end prototypes, production-ready solutions, and reusable assets, including AI-enabled applications, agentic workflows, APIs, workflow automations, and integrations Translate AI innovation concepts, including GenAI, RAG, and agentic patterns, into production-grade applications using Google Cloud technologies such as Vertex AI and Gemini Architect and implement agent-based and multi-agent systems using frameworks such as Google ADK, Agent Engine, LangGraph, or equivalent technologies Design and oversee cloud-native deployment patterns for AI applications and agent services, ideally using Cloud Run, CI/CD, observability, logging, monitoring, and secure environment configuration Architect or integrate with AI gateways and model-proxy layers supporting model routing, authentication, access controls, guardrails, policy enforcement, and usage governance Define grounding and Retrieval-Augmented Generation (RAG) patterns using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Establish AI governance patterns covering token usage, model consumption, cost management, security controls, guardrails, evaluation, and observability Build solutions with a future-proof mindset, laying a solid foundation for development and preventing the need for costly redesigns Establish and evolve engineering standards for innovation delivery, including reference architectures, reusable templates, code quality, API patterns, agent orchestration patterns, integration patterns, and secure AI development standards Partner with other technology teams to align on enterprise architecture, identity / access patterns, data access, compliance needs, private networking, and reuse opportunities Evaluate emerging technologies and developer tooling, including Digital Assets, AI-assisted development, orchestration frameworks, app frameworks, integration approaches, and AI governance technologies, and assess applicability to the client's environment Drive technical problem-solving across ambiguous domains; proactively decompose work into deliverable increments and lead execution to complete finished outcomes Provide technical mentorship and guidance to engineers, helping improve development velocity, quality, architectural consistency, and confidence across the team Contribute to innovation reporting by summarizing progress, decisions, tradeoffs, and outcomes in a clear way for senior stakeholders Champion a culture of innovation and forward thinking, educating teams about new possibilities and bring an "art of the possible" mindset to problem-solving Primary Requirements: Bachelor's degree in Computer Science, Information Technology, Artificial Intelligence, Machine Learning, or a related technical field Demonstrated hands-on experience building applications end-to-end (backend + integration + UI), including APIs and workflow automation solutions Deep hands-on Google Cloud engineering experience with demonstrated production deployment of applications and AI workloads within GCP Demonstrated production experience building AI-enabled applications using Vertex AI and Gemini; experience should extend beyond using Gemini or other GenAI tools solely as coding assistants Demonstrated hands-on experience architecting and building agentic or multi-agent systems using Google ADK, Agent Engine, LangGraph, or comparable orchestration frameworks Demonstrated experience building AI-enabled applications involving GenAI, RAG, model integration, evaluation, guardrails, and production deployment Strong experience designing AI gateway, LLM gateway, or model-proxy patterns covering model routing, authentication, security, policy enforcement, and governance Expert-level software engineering and application architecture skills with the ability to independently lead complex initiatives from concept to execution Comfortable operating in a fast-paced, ambiguous environment while managing multiple priorities and driving workplans proactively 10+ years of experience in software engineering, technology architecture, and end-to-end innovative solution delivery Desired Qualifications: Experience building innovative digital products in a regulated environment (financial services preferred) with strong security and risk-awareness Experience deploying AI solutions in highly controlled environments using technologies or patterns such as VPC Service Controls (VPC-SC), private networking, DLP, and secure service-to-service communication Strong API design skills, integration patterns, and experience building services that can be reused across multiple use cases Experience deploying agentic and AI-enabled applications using Google Cloud Run or comparable cloud-native container platforms Experience implementing grounding/RAG solutions using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Experience implementing AI safety and security controls using Model Armor or comparable technologies Experience managing token usage, model consumption, observability, and AI-related cost governance Experience with modern front-end frameworks and rapid prototyping approaches (e.g., React/Next.js or equivalent), plus strong backend development depth Strong hands-on software engineering experience using one or more modern programming languages (e.g., Java, Python, TypeScript/JavaScript, or similar), with the ability to rapidly learn and apply new technologies as needed Demonstrated success creating reusable accelerators, including templates, libraries, reference implementations, AI patterns, and governance frameworks that scale delivery across teams Experience with AWS Bedrock, Azure OpenAI, or comparable AI platforms is beneficial when combined with demonstrated hands-on Google Cloud AI deployment experience Strong communication skills with the ability to explain architecture decisions, tradeoffs, and outcomes to technical and non-technical audiences Experience building with a product-centric mindset and ability to iteratively delivery; ability to create momentum while maintaining engineering discipline Only candidates available and ready to work directly as Genesis10 employees will be considered for this position. If you have the described qualifications and are interested in this exciting opportunity, please apply! Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals. For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10: Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years Access to an experienced, caring recruiting team (more than 7 years of experience, on average) Behavioral Health Platform Medical, Dental, Vision Health Savings Account Voluntary Hospital Indemnity (Critical Illness & Accident) Voluntary Term Life Insurance 401K Sick Pay (for applicable states/municipalities) Commuter Benefits (Dallas, NYC, SF, and Illinois) For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance . click apply for full job details
09/20/2026
Full time
Job Description Job Description Genesis10 is currently seeking a Innovation Principal Engineer for a Contract-to-Hire opportunity. This position can work hybrid or can work remotely in the footprint locations of Columbus, OH; Minnetonka, MN; Atlanta, GA: Dallas, TX; Detroit, MI; Chicago, IL; Charlotte, NC; Birmingham, AL; Tupelo, MS; Pittsburgh, PA; Cleveland, OH; Akron, OH and Cincinnati, OH. Compensation: $100.00 - $110.00 per hour, W2, based on qualifications Position Overview: This position will be responsible for driving this organizations most complex innovation initiatives by designing and building end-to-end, production-grade prototypes and reusable technology assets, including AI-enabled and agentic applications within Google Cloud, while setting high engineering standards for solution architecture, application design, security, governance, and delivery acceleration across the Technology Innovation portfolio. This role requires strong hands-on full-stack engineering skills, the mindset of a startup innovator in a large organization, deep experience building AI-integrated applications end-to-end, developing APIs and user experiences, and the ability to operate independently with limited supervision in a fast-paced innovation environment. The ideal candidate will bring demonstrated production experience with Vertex AI, Gemini, agent frameworks, AI gateway/model-proxy patterns, and enterprise-grade deployment of AI workloads. Key Responsibilities: Lead the design and delivery of high-impact innovation solutions aligned with the organization's business goals and technology trends in financial services Architect and build end-to-end prototypes, production-ready solutions, and reusable assets, including AI-enabled applications, agentic workflows, APIs, workflow automations, and integrations Translate AI innovation concepts, including GenAI, RAG, and agentic patterns, into production-grade applications using Google Cloud technologies such as Vertex AI and Gemini Architect and implement agent-based and multi-agent systems using frameworks such as Google ADK, Agent Engine, LangGraph, or equivalent technologies Design and oversee cloud-native deployment patterns for AI applications and agent services, ideally using Cloud Run, CI/CD, observability, logging, monitoring, and secure environment configuration Architect or integrate with AI gateways and model-proxy layers supporting model routing, authentication, access controls, guardrails, policy enforcement, and usage governance Define grounding and Retrieval-Augmented Generation (RAG) patterns using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Establish AI governance patterns covering token usage, model consumption, cost management, security controls, guardrails, evaluation, and observability Build solutions with a future-proof mindset, laying a solid foundation for development and preventing the need for costly redesigns Establish and evolve engineering standards for innovation delivery, including reference architectures, reusable templates, code quality, API patterns, agent orchestration patterns, integration patterns, and secure AI development standards Partner with other technology teams to align on enterprise architecture, identity / access patterns, data access, compliance needs, private networking, and reuse opportunities Evaluate emerging technologies and developer tooling, including Digital Assets, AI-assisted development, orchestration frameworks, app frameworks, integration approaches, and AI governance technologies, and assess applicability to the client's environment Drive technical problem-solving across ambiguous domains; proactively decompose work into deliverable increments and lead execution to complete finished outcomes Provide technical mentorship and guidance to engineers, helping improve development velocity, quality, architectural consistency, and confidence across the team Contribute to innovation reporting by summarizing progress, decisions, tradeoffs, and outcomes in a clear way for senior stakeholders Champion a culture of innovation and forward thinking, educating teams about new possibilities and bring an "art of the possible" mindset to problem-solving Primary Requirements: Bachelor's degree in Computer Science, Information Technology, Artificial Intelligence, Machine Learning, or a related technical field Demonstrated hands-on experience building applications end-to-end (backend + integration + UI), including APIs and workflow automation solutions Deep hands-on Google Cloud engineering experience with demonstrated production deployment of applications and AI workloads within GCP Demonstrated production experience building AI-enabled applications using Vertex AI and Gemini; experience should extend beyond using Gemini or other GenAI tools solely as coding assistants Demonstrated hands-on experience architecting and building agentic or multi-agent systems using Google ADK, Agent Engine, LangGraph, or comparable orchestration frameworks Demonstrated experience building AI-enabled applications involving GenAI, RAG, model integration, evaluation, guardrails, and production deployment Strong experience designing AI gateway, LLM gateway, or model-proxy patterns covering model routing, authentication, security, policy enforcement, and governance Expert-level software engineering and application architecture skills with the ability to independently lead complex initiatives from concept to execution Comfortable operating in a fast-paced, ambiguous environment while managing multiple priorities and driving workplans proactively 10+ years of experience in software engineering, technology architecture, and end-to-end innovative solution delivery Desired Qualifications: Experience building innovative digital products in a regulated environment (financial services preferred) with strong security and risk-awareness Experience deploying AI solutions in highly controlled environments using technologies or patterns such as VPC Service Controls (VPC-SC), private networking, DLP, and secure service-to-service communication Strong API design skills, integration patterns, and experience building services that can be reused across multiple use cases Experience deploying agentic and AI-enabled applications using Google Cloud Run or comparable cloud-native container platforms Experience implementing grounding/RAG solutions using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Experience implementing AI safety and security controls using Model Armor or comparable technologies Experience managing token usage, model consumption, observability, and AI-related cost governance Experience with modern front-end frameworks and rapid prototyping approaches (e.g., React/Next.js or equivalent), plus strong backend development depth Strong hands-on software engineering experience using one or more modern programming languages (e.g., Java, Python, TypeScript/JavaScript, or similar), with the ability to rapidly learn and apply new technologies as needed Demonstrated success creating reusable accelerators, including templates, libraries, reference implementations, AI patterns, and governance frameworks that scale delivery across teams Experience with AWS Bedrock, Azure OpenAI, or comparable AI platforms is beneficial when combined with demonstrated hands-on Google Cloud AI deployment experience Strong communication skills with the ability to explain architecture decisions, tradeoffs, and outcomes to technical and non-technical audiences Experience building with a product-centric mindset and ability to iteratively delivery; ability to create momentum while maintaining engineering discipline Only candidates available and ready to work directly as Genesis10 employees will be considered for this position. If you have the described qualifications and are interested in this exciting opportunity, please apply! Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals. For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10: Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years Access to an experienced, caring recruiting team (more than 7 years of experience, on average) Behavioral Health Platform Medical, Dental, Vision Health Savings Account Voluntary Hospital Indemnity (Critical Illness & Accident) Voluntary Term Life Insurance 401K Sick Pay (for applicable states/municipalities) Commuter Benefits (Dallas, NYC, SF, and Illinois) For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance . click apply for full job details
Job Description Job Description About Company: Papigen is a fast-growing global technology services company delivering innovative, enterprise-grade digital solutions across industries. We specialize in technology transformation, enterprise modernization, and advanced capabilities including Cloud, Data Engineering, Artificial Intelligence, and Digital Platforms. With deep domain expertise in financial services and complex enterprise environments, Papigen partners with organizations to modernize legacy systems, enable data-driven decision-making, and build scalable, secure, and high-performing applications. Our services span consulting, product engineering, data analytics, DevOps, and intelligent automation, helping clients accelerate innovation while ensuring operational excellence and regulatory compliance. We adopt a client-centric approach that combines strategic thinking with hands-on engineering execution. Our teams work collaboratively across global delivery centers, leveraging Agile and SAFe methodologies to deliver high-quality solutions in fast-paced environments. At Papigen, we foster a culture of innovation, continuous learning, and collaboration, providing professionals with opportunities to work on cutting-edge technologies, solve complex business challenges, and contribute to impactful digital transformation initiatives. Job Title: Senior Developer III (PowerApps, .NET Core & Azure) Job Location: Washington D.C (Hybrid) Role Summary: We are seeking a Senior Developer III to lead the design and development of a PowerApps-based Syndications Management Platform. The platform will serve as a centralized solution for project lifecycle management, fund allocation automation, investor interactions, workflow orchestration, and real-time reporting. The ideal candidate will possess strong expertise in Microsoft Power Platform, .NET Core, Azure, and enterprise integration architectures, with the ability to translate business requirements into scalable technical solutions. Scope of Work & Key Responsibilities: Design and implement technical architecture for a unified syndications management platform. Develop workflow solutions for pipeline-to-portfolio project tracking and management. Build and automate fund allocation models based on eligibility criteria, allocation rules, and audit requirements. Design real-time synchronization and integration between enterprise systems and applications. Implement stage management workflows, notifications, alerts, and investor engagement capabilities. Develop scalable solutions using PowerApps, .NET Core, Azure services, and Dataverse. Integrate reporting and analytics capabilities using Power BI and enterprise data sources. Establish security, data access controls, and compliance standards for platform operations. Support event-driven automation and workflow orchestration across business processes. Optimize application performance, scalability, reliability, and maintainability. Collaborate with product owners, architects, developers, and business stakeholders throughout the project lifecycle. Participate in Agile ceremonies, technical reviews, and solution planning activities. Required Skills & Experience: Bachelor's degree in Computer Science, Information Technology, Engineering, or related discipline (or equivalent practical experience). 8+ years of experience with Microsoft Power Apps and business process automation solutions. Proven experience designing and delivering enterprise applications using .NET Core and React JS. Strong expertise in Microsoft Power Platform, Dataverse, and workflow automation. Experience working with Microsoft Azure services and cloud-native architectures. Strong knowledge of enterprise systems integration, API connectivity, and secure application design. Experience with Power BI, SQL, reporting frameworks, and data validation techniques. Strong understanding of financial systems, CRM platforms, SAP integrations, and loan servicing processes. Experience with Dataverse data modeling and application architecture. Knowledge of event-driven architectures and automation frameworks. Experience working within Agile delivery environments. Strong communication, stakeholder management, and technical documentation skills. Preferred Skills: Experience developing financial modeling or allocation management systems. Familiarity with AI and Copilot capabilities within the Power Platform ecosystem. Experience designing enterprise reporting and dashboard solutions. Exposure to Azure AI services and intelligent automation technologies. Why Join Us: Opportunity to lead enterprise-scale cloud and AI transformation initiatives. Work on cutting-edge Agentic AI architectures and multi-cloud platforms. Collaborative environment with global teams and high-impact projects. Exposure to modern DevSecOps, platform engineering, and AI-driven solutions. Strong focus on innovation, leadership growth, and continuous learning in emerging technologies.
09/20/2026
Full time
Job Description Job Description About Company: Papigen is a fast-growing global technology services company delivering innovative, enterprise-grade digital solutions across industries. We specialize in technology transformation, enterprise modernization, and advanced capabilities including Cloud, Data Engineering, Artificial Intelligence, and Digital Platforms. With deep domain expertise in financial services and complex enterprise environments, Papigen partners with organizations to modernize legacy systems, enable data-driven decision-making, and build scalable, secure, and high-performing applications. Our services span consulting, product engineering, data analytics, DevOps, and intelligent automation, helping clients accelerate innovation while ensuring operational excellence and regulatory compliance. We adopt a client-centric approach that combines strategic thinking with hands-on engineering execution. Our teams work collaboratively across global delivery centers, leveraging Agile and SAFe methodologies to deliver high-quality solutions in fast-paced environments. At Papigen, we foster a culture of innovation, continuous learning, and collaboration, providing professionals with opportunities to work on cutting-edge technologies, solve complex business challenges, and contribute to impactful digital transformation initiatives. Job Title: Senior Developer III (PowerApps, .NET Core & Azure) Job Location: Washington D.C (Hybrid) Role Summary: We are seeking a Senior Developer III to lead the design and development of a PowerApps-based Syndications Management Platform. The platform will serve as a centralized solution for project lifecycle management, fund allocation automation, investor interactions, workflow orchestration, and real-time reporting. The ideal candidate will possess strong expertise in Microsoft Power Platform, .NET Core, Azure, and enterprise integration architectures, with the ability to translate business requirements into scalable technical solutions. Scope of Work & Key Responsibilities: Design and implement technical architecture for a unified syndications management platform. Develop workflow solutions for pipeline-to-portfolio project tracking and management. Build and automate fund allocation models based on eligibility criteria, allocation rules, and audit requirements. Design real-time synchronization and integration between enterprise systems and applications. Implement stage management workflows, notifications, alerts, and investor engagement capabilities. Develop scalable solutions using PowerApps, .NET Core, Azure services, and Dataverse. Integrate reporting and analytics capabilities using Power BI and enterprise data sources. Establish security, data access controls, and compliance standards for platform operations. Support event-driven automation and workflow orchestration across business processes. Optimize application performance, scalability, reliability, and maintainability. Collaborate with product owners, architects, developers, and business stakeholders throughout the project lifecycle. Participate in Agile ceremonies, technical reviews, and solution planning activities. Required Skills & Experience: Bachelor's degree in Computer Science, Information Technology, Engineering, or related discipline (or equivalent practical experience). 8+ years of experience with Microsoft Power Apps and business process automation solutions. Proven experience designing and delivering enterprise applications using .NET Core and React JS. Strong expertise in Microsoft Power Platform, Dataverse, and workflow automation. Experience working with Microsoft Azure services and cloud-native architectures. Strong knowledge of enterprise systems integration, API connectivity, and secure application design. Experience with Power BI, SQL, reporting frameworks, and data validation techniques. Strong understanding of financial systems, CRM platforms, SAP integrations, and loan servicing processes. Experience with Dataverse data modeling and application architecture. Knowledge of event-driven architectures and automation frameworks. Experience working within Agile delivery environments. Strong communication, stakeholder management, and technical documentation skills. Preferred Skills: Experience developing financial modeling or allocation management systems. Familiarity with AI and Copilot capabilities within the Power Platform ecosystem. Experience designing enterprise reporting and dashboard solutions. Exposure to Azure AI services and intelligent automation technologies. Why Join Us: Opportunity to lead enterprise-scale cloud and AI transformation initiatives. Work on cutting-edge Agentic AI architectures and multi-cloud platforms. Collaborative environment with global teams and high-impact projects. Exposure to modern DevSecOps, platform engineering, and AI-driven solutions. Strong focus on innovation, leadership growth, and continuous learning in emerging technologies.
Job Description Job Description Genesis10 is currently seeking a Innovation Principal Engineer for a Contract-to-Hire opportunity. This position can work hybrid or can work remotely in the footprint locations of Columbus, OH; Minnetonka, MN; Atlanta, GA: Dallas, TX; Detroit, MI; Chicago, IL; Charlotte, NC; Birmingham, AL; Tupelo, MS; Pittsburgh, PA; Cleveland, OH; Akron, OH and Cincinnati, OH. Compensation: $100.00 - $110.00 per hour, W2, based on qualifications Position Overview: This position will be responsible for driving this organizations most complex innovation initiatives by designing and building end-to-end, production-grade prototypes and reusable technology assets, including AI-enabled and agentic applications within Google Cloud, while setting high engineering standards for solution architecture, application design, security, governance, and delivery acceleration across the Technology Innovation portfolio. This role requires strong hands-on full-stack engineering skills, the mindset of a startup innovator in a large organization, deep experience building AI-integrated applications end-to-end, developing APIs and user experiences, and the ability to operate independently with limited supervision in a fast-paced innovation environment. The ideal candidate will bring demonstrated production experience with Vertex AI, Gemini, agent frameworks, AI gateway/model-proxy patterns, and enterprise-grade deployment of AI workloads. Key Responsibilities: Lead the design and delivery of high-impact innovation solutions aligned with the organization's business goals and technology trends in financial services Architect and build end-to-end prototypes, production-ready solutions, and reusable assets, including AI-enabled applications, agentic workflows, APIs, workflow automations, and integrations Translate AI innovation concepts, including GenAI, RAG, and agentic patterns, into production-grade applications using Google Cloud technologies such as Vertex AI and Gemini Architect and implement agent-based and multi-agent systems using frameworks such as Google ADK, Agent Engine, LangGraph, or equivalent technologies Design and oversee cloud-native deployment patterns for AI applications and agent services, ideally using Cloud Run, CI/CD, observability, logging, monitoring, and secure environment configuration Architect or integrate with AI gateways and model-proxy layers supporting model routing, authentication, access controls, guardrails, policy enforcement, and usage governance Define grounding and Retrieval-Augmented Generation (RAG) patterns using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Establish AI governance patterns covering token usage, model consumption, cost management, security controls, guardrails, evaluation, and observability Build solutions with a future-proof mindset, laying a solid foundation for development and preventing the need for costly redesigns Establish and evolve engineering standards for innovation delivery, including reference architectures, reusable templates, code quality, API patterns, agent orchestration patterns, integration patterns, and secure AI development standards Partner with other technology teams to align on enterprise architecture, identity / access patterns, data access, compliance needs, private networking, and reuse opportunities Evaluate emerging technologies and developer tooling, including Digital Assets, AI-assisted development, orchestration frameworks, app frameworks, integration approaches, and AI governance technologies, and assess applicability to the client's environment Drive technical problem-solving across ambiguous domains; proactively decompose work into deliverable increments and lead execution to complete finished outcomes Provide technical mentorship and guidance to engineers, helping improve development velocity, quality, architectural consistency, and confidence across the team Contribute to innovation reporting by summarizing progress, decisions, tradeoffs, and outcomes in a clear way for senior stakeholders Champion a culture of innovation and forward thinking, educating teams about new possibilities and bring an "art of the possible" mindset to problem-solving Primary Requirements: Bachelor's degree in Computer Science, Information Technology, Artificial Intelligence, Machine Learning, or a related technical field Demonstrated hands-on experience building applications end-to-end (backend + integration + UI), including APIs and workflow automation solutions Deep hands-on Google Cloud engineering experience with demonstrated production deployment of applications and AI workloads within GCP Demonstrated production experience building AI-enabled applications using Vertex AI and Gemini; experience should extend beyond using Gemini or other GenAI tools solely as coding assistants Demonstrated hands-on experience architecting and building agentic or multi-agent systems using Google ADK, Agent Engine, LangGraph, or comparable orchestration frameworks Demonstrated experience building AI-enabled applications involving GenAI, RAG, model integration, evaluation, guardrails, and production deployment Strong experience designing AI gateway, LLM gateway, or model-proxy patterns covering model routing, authentication, security, policy enforcement, and governance Expert-level software engineering and application architecture skills with the ability to independently lead complex initiatives from concept to execution Comfortable operating in a fast-paced, ambiguous environment while managing multiple priorities and driving workplans proactively 10+ years of experience in software engineering, technology architecture, and end-to-end innovative solution delivery Desired Qualifications: Experience building innovative digital products in a regulated environment (financial services preferred) with strong security and risk-awareness Experience deploying AI solutions in highly controlled environments using technologies or patterns such as VPC Service Controls (VPC-SC), private networking, DLP, and secure service-to-service communication Strong API design skills, integration patterns, and experience building services that can be reused across multiple use cases Experience deploying agentic and AI-enabled applications using Google Cloud Run or comparable cloud-native container platforms Experience implementing grounding/RAG solutions using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Experience implementing AI safety and security controls using Model Armor or comparable technologies Experience managing token usage, model consumption, observability, and AI-related cost governance Experience with modern front-end frameworks and rapid prototyping approaches (e.g., React/Next.js or equivalent), plus strong backend development depth Strong hands-on software engineering experience using one or more modern programming languages (e.g., Java, Python, TypeScript/JavaScript, or similar), with the ability to rapidly learn and apply new technologies as needed Demonstrated success creating reusable accelerators, including templates, libraries, reference implementations, AI patterns, and governance frameworks that scale delivery across teams Experience with AWS Bedrock, Azure OpenAI, or comparable AI platforms is beneficial when combined with demonstrated hands-on Google Cloud AI deployment experience Strong communication skills with the ability to explain architecture decisions, tradeoffs, and outcomes to technical and non-technical audiences Experience building with a product-centric mindset and ability to iteratively delivery; ability to create momentum while maintaining engineering discipline Only candidates available and ready to work directly as Genesis10 employees will be considered for this position. If you have the described qualifications and are interested in this exciting opportunity, please apply! Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals. For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10: Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years Access to an experienced, caring recruiting team (more than 7 years of experience, on average) Behavioral Health Platform Medical, Dental, Vision Health Savings Account Voluntary Hospital Indemnity (Critical Illness & Accident) Voluntary Term Life Insurance 401K Sick Pay (for applicable states/municipalities) Commuter Benefits (Dallas, NYC, SF, and Illinois) For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance . click apply for full job details
09/20/2026
Full time
Job Description Job Description Genesis10 is currently seeking a Innovation Principal Engineer for a Contract-to-Hire opportunity. This position can work hybrid or can work remotely in the footprint locations of Columbus, OH; Minnetonka, MN; Atlanta, GA: Dallas, TX; Detroit, MI; Chicago, IL; Charlotte, NC; Birmingham, AL; Tupelo, MS; Pittsburgh, PA; Cleveland, OH; Akron, OH and Cincinnati, OH. Compensation: $100.00 - $110.00 per hour, W2, based on qualifications Position Overview: This position will be responsible for driving this organizations most complex innovation initiatives by designing and building end-to-end, production-grade prototypes and reusable technology assets, including AI-enabled and agentic applications within Google Cloud, while setting high engineering standards for solution architecture, application design, security, governance, and delivery acceleration across the Technology Innovation portfolio. This role requires strong hands-on full-stack engineering skills, the mindset of a startup innovator in a large organization, deep experience building AI-integrated applications end-to-end, developing APIs and user experiences, and the ability to operate independently with limited supervision in a fast-paced innovation environment. The ideal candidate will bring demonstrated production experience with Vertex AI, Gemini, agent frameworks, AI gateway/model-proxy patterns, and enterprise-grade deployment of AI workloads. Key Responsibilities: Lead the design and delivery of high-impact innovation solutions aligned with the organization's business goals and technology trends in financial services Architect and build end-to-end prototypes, production-ready solutions, and reusable assets, including AI-enabled applications, agentic workflows, APIs, workflow automations, and integrations Translate AI innovation concepts, including GenAI, RAG, and agentic patterns, into production-grade applications using Google Cloud technologies such as Vertex AI and Gemini Architect and implement agent-based and multi-agent systems using frameworks such as Google ADK, Agent Engine, LangGraph, or equivalent technologies Design and oversee cloud-native deployment patterns for AI applications and agent services, ideally using Cloud Run, CI/CD, observability, logging, monitoring, and secure environment configuration Architect or integrate with AI gateways and model-proxy layers supporting model routing, authentication, access controls, guardrails, policy enforcement, and usage governance Define grounding and Retrieval-Augmented Generation (RAG) patterns using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Establish AI governance patterns covering token usage, model consumption, cost management, security controls, guardrails, evaluation, and observability Build solutions with a future-proof mindset, laying a solid foundation for development and preventing the need for costly redesigns Establish and evolve engineering standards for innovation delivery, including reference architectures, reusable templates, code quality, API patterns, agent orchestration patterns, integration patterns, and secure AI development standards Partner with other technology teams to align on enterprise architecture, identity / access patterns, data access, compliance needs, private networking, and reuse opportunities Evaluate emerging technologies and developer tooling, including Digital Assets, AI-assisted development, orchestration frameworks, app frameworks, integration approaches, and AI governance technologies, and assess applicability to the client's environment Drive technical problem-solving across ambiguous domains; proactively decompose work into deliverable increments and lead execution to complete finished outcomes Provide technical mentorship and guidance to engineers, helping improve development velocity, quality, architectural consistency, and confidence across the team Contribute to innovation reporting by summarizing progress, decisions, tradeoffs, and outcomes in a clear way for senior stakeholders Champion a culture of innovation and forward thinking, educating teams about new possibilities and bring an "art of the possible" mindset to problem-solving Primary Requirements: Bachelor's degree in Computer Science, Information Technology, Artificial Intelligence, Machine Learning, or a related technical field Demonstrated hands-on experience building applications end-to-end (backend + integration + UI), including APIs and workflow automation solutions Deep hands-on Google Cloud engineering experience with demonstrated production deployment of applications and AI workloads within GCP Demonstrated production experience building AI-enabled applications using Vertex AI and Gemini; experience should extend beyond using Gemini or other GenAI tools solely as coding assistants Demonstrated hands-on experience architecting and building agentic or multi-agent systems using Google ADK, Agent Engine, LangGraph, or comparable orchestration frameworks Demonstrated experience building AI-enabled applications involving GenAI, RAG, model integration, evaluation, guardrails, and production deployment Strong experience designing AI gateway, LLM gateway, or model-proxy patterns covering model routing, authentication, security, policy enforcement, and governance Expert-level software engineering and application architecture skills with the ability to independently lead complex initiatives from concept to execution Comfortable operating in a fast-paced, ambiguous environment while managing multiple priorities and driving workplans proactively 10+ years of experience in software engineering, technology architecture, and end-to-end innovative solution delivery Desired Qualifications: Experience building innovative digital products in a regulated environment (financial services preferred) with strong security and risk-awareness Experience deploying AI solutions in highly controlled environments using technologies or patterns such as VPC Service Controls (VPC-SC), private networking, DLP, and secure service-to-service communication Strong API design skills, integration patterns, and experience building services that can be reused across multiple use cases Experience deploying agentic and AI-enabled applications using Google Cloud Run or comparable cloud-native container platforms Experience implementing grounding/RAG solutions using Vertex AI Search / Discovery Engine or comparable enterprise retrieval technologies Experience implementing AI safety and security controls using Model Armor or comparable technologies Experience managing token usage, model consumption, observability, and AI-related cost governance Experience with modern front-end frameworks and rapid prototyping approaches (e.g., React/Next.js or equivalent), plus strong backend development depth Strong hands-on software engineering experience using one or more modern programming languages (e.g., Java, Python, TypeScript/JavaScript, or similar), with the ability to rapidly learn and apply new technologies as needed Demonstrated success creating reusable accelerators, including templates, libraries, reference implementations, AI patterns, and governance frameworks that scale delivery across teams Experience with AWS Bedrock, Azure OpenAI, or comparable AI platforms is beneficial when combined with demonstrated hands-on Google Cloud AI deployment experience Strong communication skills with the ability to explain architecture decisions, tradeoffs, and outcomes to technical and non-technical audiences Experience building with a product-centric mindset and ability to iteratively delivery; ability to create momentum while maintaining engineering discipline Only candidates available and ready to work directly as Genesis10 employees will be considered for this position. If you have the described qualifications and are interested in this exciting opportunity, please apply! Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals. For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10: Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years Access to an experienced, caring recruiting team (more than 7 years of experience, on average) Behavioral Health Platform Medical, Dental, Vision Health Savings Account Voluntary Hospital Indemnity (Critical Illness & Accident) Voluntary Term Life Insurance 401K Sick Pay (for applicable states/municipalities) Commuter Benefits (Dallas, NYC, SF, and Illinois) For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance . click apply for full job details
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
09/20/2026
Full time
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
Charlie Health Engineering, Product & Design
New York, New York
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
09/20/2026
Full time
Job Description Job Description Why Charlie Health? Millions of people across the country are navigating mental health conditions, substance use disorders, and eating disorders, but too often, they're met with barriers to care. From limited local options and long wait times to treatment that lacks personalization, behavioral healthcare can leave people feeling unseen and unsupported. Charlie Health exists to change that. Our mission is to connect the world to life-saving behavioral health treatment. We deliver personalized, virtual care rooted in connection-between clients and clinicians, care teams, loved ones, and the communities that support them. By focusing on people with complex needs, we're expanding access to meaningful care and driving better outcomes from the comfort of home. As a rapidly growing organization, we're reaching more communities every day and building a team that's redefining what behavioral health treatment can look like. If you're ready to use your skills to drive lasting change and help more people access the care they deserve, we'd love to meet you. About the Role Core Platform is a platform team. Our customers are Charlie Health's other engineering teams, and our job is to make them faster, safer, and more consistent. We don't ship product features. We build the patterns, services, and guardrails that other engineering teams depend on. As the Principal Platform Engineer on Core Platform, you will be the technical lead for the team building the platform underneath everything Charlie Health ships next: a federated GraphQL front door that fronts every client surface, the event substrate that carries every signal, the identity layer that authenticates every human and every AI agent, and the real-time world model state layer that gives our services and agents a live picture of care. You will be foundational in a replatforming effort, and the primitives you design become the platform's paved roads. This is a technical leadership role without direct reports. You own the team's technical direction, roadmap, and delivery. You will set direction for a team of Staff and Senior Platform Engineers, make the hardest design calls of the replatforming, and stay hands-on in the work. We build with AI as the default. Specs are the primary artifact, agents draft most of the implementation, and the quality gates this team owns are what make that safe in a HIPAA-regulated clinical business. Responsibilities Own the team's roadmap and intake. Balance incoming platform requests and near-term developer pain against long-term architectural investment, with support from the Director of Platform Engineering and the CTO Set the technical direction for the Core Platform team and break the north star architecture into projects the team ships quarter by quarter Run the team's delivery: planning, refinement, and the sprint cadence, with epics scoped to clear milestones and target dates Architect the federated GraphQL supergraph and govern its schema, including subgraph onboarding standards, contract enforcement, and caching and persisted query posture Build authentication and authorization for the platform, including agent identity: scoped, short-lived, audited credentials for non-human principals Design, build, and operate the centralized event substrate, including schema governance, dead-letter conventions, and the producer and consumer patterns that make async the default integration path between services Guide the technical execution of our strangler-fig migration through to cutover, working with product squads as they move onto the new platform Deliver the real-time world model state layer: the stream processing that turns events into state, and the live, low-latency state graph that services and agents query Facilitate the Architecture Advice Guild rituals and provide technical direction on architecture decision records for cross-team decisions and standards Raise the bar for spec writing and agent orchestration on the team, and review specs and agent output with the same rigor you would apply to a senior engineer's work Keep the platform services the team operates reliable: SLOs, monitor-gated deploys, incident leadership, and a seat in the on-call rotation Drive the technical evaluation for procurement and renewals of our critical platform vendors: evaluations, proofs of concept, usage and cost data, and build versus buy recommendations, partnering with the Director who owns the commercial side Mentor Staff and Senior engineers toward larger scope Requirements 10+ years of software engineering experience, including ownership of platform or distributed systems architecture at company scale. You have designed, operated, and evolved systems that other teams build on Hands-on experience with GraphQL federation at scale and with event streaming systems such as Kafka, covering design, schema governance, and production operation rather than just consumption Experience with skills, MCPs, harnesses, agentic tooling ecosystems, or agent orchestration for software development A track record of technical leadership without people authority. You have aligned multiple teams behind architectural outcomes and developed senior engineers Experience owning a team roadmap: intake, prioritization, and delivery planning alongside the architecture work You build with AI agents as part of your daily work. You write specs agents can implement, run agents for real delivery, and know where agent output needs human judgment Proficiency in the stack we run: TypeScript and/or Python services, Postgres, and AWS managed primitives Comfort operating production critical-path systems, including on-call, in a regulated environment Clear, direct communication with engineers, leadership, clinicians, and compliance This role requires 4 days per week in our NYC office. Nice to Haves Real-time stream processing (Flink, Spark, KsqlDB, or similar) and graph databases (Memgraph, Neo4j, Amazon Neptune, Dgraph, or similar) Healthcare data standards (FHIR, HL7v2) and engineering in a HIPAA-regulated environment Time as a people manager for senior or staff level engineers A strangler-fig migration or large platform cutover you have led Prior work on a platform engineering team Fluency with domain-driven design or event storming Benefits Charlie Health is pleased to offer comprehensive benefits to all full-time, exempt employees. Read more about our benefits here. The total target base compensation for this role will be between $225,000 and $325,000 per year at the commencement of employment. Please note, pay will be determined on an individualized basis and will be impacted by location, experience, expertise, internal pay equity, and other relevant business considerations. Further, cash compensation is only part of the total compensation package, which, depending on the position, may include stock options and other Charlie Health-sponsored benefits. Our Values Connection: Care deeply & inspire hope. Congruence: Stay curious & heed the evidence. Commitment: Act with urgency & don't give up. Please do not call our public clinical admissions line in regard to this or any other job posting. Please be cautious of potential recruitment fraud. If you are interested in exploring opportunities at Charlie Health, please go directly to our Careers Page: -openings. Charlie Health will never ask you to pay a fee or download software as part of the interview process with our company. In addition, Charlie Health will not ask for your personal banking information until you have signed an offer of employment and completed onboarding paperwork that is provided by our People Operations team. All communications with Charlie Health Talent and People Operations professionals will only be sent email addresses. Legitimate emails will never originate from or other commercial email services. Recruiting agencies, please do not submit unsolicited referrals for this or any open role. We have a roster of agencies with whom we partner, and we will not pay any fee associated with unsolicited referrals. At Charlie Health, we value being an Equal Opportunity Employer. We strive to cultivate an environment where individuals can be their authentic selves. Being an Equal Opportunity Employer means every member of our team feels as though they are supported and belong. We value diverse perspectives to help us provide essential mental health and substance use disorder treatments to all young people. Charlie Health applicants are assessed solely on their qualifications for the role, without regard to disability or need for accommodation. By clicking "Submit application" below, you agree to Charlie Health's Privacy Policy and Terms of Service. By submitting your application, you agree to receive SMS messages from Charlie Health regarding your application. Message and data rates may apply. Message frequency varies. You can reply STOP to opt out at any time. For help, reply HELP.
Job Description An exciting career awaits you At MPC, we're committed to being a great place to work - one that welcomes new ideas, encourages diverse perspectives, develops our people, and fosters a collaborative team environment. Position Summary The GRC Automation & Continuous Controls Monitoring (CCM) Senior Cybersecurity Engineer is a strategic and technical role, responsible or helping turn GRC into a trust engine for the business: reducing audit friction, improving business risk exposure visibility, accelerating evidence readiness, and enabling stronger operational confidence. The role serves as the technical focal for GRC automation, developing automated evidence collection, control testing, risk intelligence, and AI-enabled governance capabilities across cloud, on-premises, identity, operational technology (OT), and security platforms. The role collaborates closely with Cyber Fusion, Enterprise Architecture, and Cyber Engineering, to design and build deterministic control-testing logic where deterministic approaches are enough; and AI agents and LLM-backed workflows where reasoning, summarization, or judgment is required, such as evidence-to-control mapping, attestation drafting, drift detection, and audit-package assembly. This position belongs to a family of jobs with increasing responsibility, competency, and skill level. Actual position title and pay grade will be based on the selected candidate's experience and qualifications. Key Responsibilities Conducts detailed analyses on changes to cybersecurity solutions and its relationship to internal and external systems to assess control effectiveness and cybersecurity risk. Resolves complex multi-functional technical issues. Leverages cybersecurity assessments, standards, control testing methodologies, and compliance frameworks to ensure compliance across security systems. Improves the efficiency and effectiveness of Security solutions, governance processes, automated controls, and monitoring capabilities. Analyzes existing processes and procedures and leads efforts for implementing improvements, automation opportunities, or remediation activities. Responsible for development and submission of Standard Operating Procedures. Analyzes business impacting events, performs initial investigation, and evaluates control performance, exceptions, and risk indicators through continuous monitoring activities. Investigates and analyzes the nature and scope of cyber incidents, control failures, and compliance exceptions. Assists in the development of risk mitigation and remediation plans to ensure regulatory and internal compliance. compliance. Leads implementation of global security initiatives, policies, compliance requirements, and continuous control monitoring practices. Collects, validates, and reports security metrics, control performance results, and remediation efforts associated with them. Manages cyber security-related consulting, guidance, and support to customers and stakeholders. Translates security principles to assist configuration teams with incorporating security and compliance requirements into build and configuration processes. Monitors emerging IT/OT, cybersecurity, automation, and artificial intelligence technologies as well as their impact on the security, risk, and compliance landscape. Education and Experience Bachelor's Degree in Information Technology, related field or equivalent experience. 5+ years of relevant experience required Experience designing, implementing, and scaling GRC, CCM, or compliance automation solutions within a regulated environment required Experience in Python or comparable automation technologies, proficient in developing, integrating, and supporting automated workflows and REST API-based data integrations across multiple enterprise systems required. Hands-on experience integrating security-control data sources into a GRC / CCM evidence pipeline for continuous control testing required. This includes normalizing API, telemetry, configuration, vulnerability, identity, ticketing, and assessment data into control-level evidence, exception logic, ownership, frequency, and audit-ready records across at least three domains such as CNAPP / CSPM, SIEM / security data lake / XDR, CTEM / VM / ASPM, DSPM, ITSM, IAM, GRC / CCM, or AI/agent governance required. Experience building LLM-backed agentic workflows on Azure AI Foundry, GitHub, or open-source frameworks preferred. Familiarity with NIST AI RMF, the OWASP LLM Top 10, or comparable AI risk frameworks preferred. Skills Adaptability - Maintaining effectiveness when experiencing major changes in work responsibilities or environment (e.g., people, processes, structure, or culture); adjusting effectively to change by exploring the benefits, trying new approaches, and collaborating with others to make the change successful. AI Fundamentals - Understanding of core AI concepts and methods, ability to apply AI to job-relevant use cases and capacity to contribute to organizational AI reimagination. Change Management - Change Management refers to a systematic approach for defining and implementing procedures and/or technologies to deal with changes in the environment. It can mean adapting to change, controlling change and/or effecting change. Authentic Communicator - Expresses ideas and information, both verbally and in writing, clearly and credibly. Listens to understand and fosters constructive dialogue. Cybersecurity Risk Management - The process of developing cyber risk assessment and treatment techniques that can effectively pre-empt and identify significant security loopholes and weaknesses, demonstrating the business risks associated with these loopholes and providing risk treatment and prioritization strategies to effectively address the cyber-related risks, threats and vulnerabilities, ensuring appropriate levels of protection, confidentiality, integrity and privacy in alignment with the security framework. General Programming - Applies a computer language to communicate with computers using a set of instructions and to automate the execution of tasks. Intrusion Detection - The use of security analytics, including the outputs from intelligence analysis, predictive research and root cause analysis in order to search for and detect potential breaches or identify recognized indicators and warnings. Also, monitoring and collating external vulnerability reports for organizational relevance, ensuring that relevant vulnerabilities are rectified through formal change processes. Penetration Testing - The practice of testing a computer system, network or web application to find security vulnerabilities that an attacker could exploit. Penetration testing can be automated with software applications or performed manually. Relationship Management - Relationship Management is the conscious aim to develop and manage long-term and/or trusting relationships with internal or external customers, distributors, suppliers, or other parties in an environment which can include marketing, selling, servicing and other areas where a relationship is crucial to on-going success. At a senior level, it includes C-level relationships with senior management. Security Controls - Manages and maintains an information system that focuses on the management of risk and the management of information systems security. Security Governance - The process of developing and disseminating corporate security policies, frameworks and guidelines to ensure that day-to-day business operations are guarded and well protected against risks, threats and vulnerabilities. Security Information & Event Management (SIEM) - A set of tools and services offering real-time visibility across an organization's information security systems, and event log management that consolidates data from numerous sources. Security Policy Management - The process of identifying, implementing, and managing the rules and procedures that all individuals must follow when accessing and using an organization's IT assets and resources. Threat Analysis - Monitor intelligence-gathering and anticipate potential threats to an IT/OT systems proactively. This involves the pre-emptive analysis of potential perpetrators, anomalous activities and evidence-based knowledge and inferences on perpetrators' motivations and tactics. Threat Hunting - Searches through networks, endpoints, and datasets to detect and isolate cyber threats that evade existing security solutions. Vulnerability Management - The process of defining, identifying, classifying and prioritizing vulnerabilities in computer systems, applications and network infrastructures and providing the organization with the necessary knowledge, awareness and risk background to understand the threats to its business MINIMUM QUALIFICATIONS: Bachelor's Degree in Information Technology, related field or equivalent experience. Professional certification, e.g. Security+, Network+, OSCP, GIAC, CEH preferred. 5+ years of relevant experience required As an energy industry leader, our career opportunities fuel personal and professional growth. Location: San Antonio, Texas Additional locations: Findlay, Ohio, Houston, Texas Job Requisition ID: Location Address: 19100 Ridgewood Pkwy Education: Employee Group: Full time Employee Subgroup: . click apply for full job details
09/19/2026
Full time
Job Description An exciting career awaits you At MPC, we're committed to being a great place to work - one that welcomes new ideas, encourages diverse perspectives, develops our people, and fosters a collaborative team environment. Position Summary The GRC Automation & Continuous Controls Monitoring (CCM) Senior Cybersecurity Engineer is a strategic and technical role, responsible or helping turn GRC into a trust engine for the business: reducing audit friction, improving business risk exposure visibility, accelerating evidence readiness, and enabling stronger operational confidence. The role serves as the technical focal for GRC automation, developing automated evidence collection, control testing, risk intelligence, and AI-enabled governance capabilities across cloud, on-premises, identity, operational technology (OT), and security platforms. The role collaborates closely with Cyber Fusion, Enterprise Architecture, and Cyber Engineering, to design and build deterministic control-testing logic where deterministic approaches are enough; and AI agents and LLM-backed workflows where reasoning, summarization, or judgment is required, such as evidence-to-control mapping, attestation drafting, drift detection, and audit-package assembly. This position belongs to a family of jobs with increasing responsibility, competency, and skill level. Actual position title and pay grade will be based on the selected candidate's experience and qualifications. Key Responsibilities Conducts detailed analyses on changes to cybersecurity solutions and its relationship to internal and external systems to assess control effectiveness and cybersecurity risk. Resolves complex multi-functional technical issues. Leverages cybersecurity assessments, standards, control testing methodologies, and compliance frameworks to ensure compliance across security systems. Improves the efficiency and effectiveness of Security solutions, governance processes, automated controls, and monitoring capabilities. Analyzes existing processes and procedures and leads efforts for implementing improvements, automation opportunities, or remediation activities. Responsible for development and submission of Standard Operating Procedures. Analyzes business impacting events, performs initial investigation, and evaluates control performance, exceptions, and risk indicators through continuous monitoring activities. Investigates and analyzes the nature and scope of cyber incidents, control failures, and compliance exceptions. Assists in the development of risk mitigation and remediation plans to ensure regulatory and internal compliance. compliance. Leads implementation of global security initiatives, policies, compliance requirements, and continuous control monitoring practices. Collects, validates, and reports security metrics, control performance results, and remediation efforts associated with them. Manages cyber security-related consulting, guidance, and support to customers and stakeholders. Translates security principles to assist configuration teams with incorporating security and compliance requirements into build and configuration processes. Monitors emerging IT/OT, cybersecurity, automation, and artificial intelligence technologies as well as their impact on the security, risk, and compliance landscape. Education and Experience Bachelor's Degree in Information Technology, related field or equivalent experience. 5+ years of relevant experience required Experience designing, implementing, and scaling GRC, CCM, or compliance automation solutions within a regulated environment required Experience in Python or comparable automation technologies, proficient in developing, integrating, and supporting automated workflows and REST API-based data integrations across multiple enterprise systems required. Hands-on experience integrating security-control data sources into a GRC / CCM evidence pipeline for continuous control testing required. This includes normalizing API, telemetry, configuration, vulnerability, identity, ticketing, and assessment data into control-level evidence, exception logic, ownership, frequency, and audit-ready records across at least three domains such as CNAPP / CSPM, SIEM / security data lake / XDR, CTEM / VM / ASPM, DSPM, ITSM, IAM, GRC / CCM, or AI/agent governance required. Experience building LLM-backed agentic workflows on Azure AI Foundry, GitHub, or open-source frameworks preferred. Familiarity with NIST AI RMF, the OWASP LLM Top 10, or comparable AI risk frameworks preferred. Skills Adaptability - Maintaining effectiveness when experiencing major changes in work responsibilities or environment (e.g., people, processes, structure, or culture); adjusting effectively to change by exploring the benefits, trying new approaches, and collaborating with others to make the change successful. AI Fundamentals - Understanding of core AI concepts and methods, ability to apply AI to job-relevant use cases and capacity to contribute to organizational AI reimagination. Change Management - Change Management refers to a systematic approach for defining and implementing procedures and/or technologies to deal with changes in the environment. It can mean adapting to change, controlling change and/or effecting change. Authentic Communicator - Expresses ideas and information, both verbally and in writing, clearly and credibly. Listens to understand and fosters constructive dialogue. Cybersecurity Risk Management - The process of developing cyber risk assessment and treatment techniques that can effectively pre-empt and identify significant security loopholes and weaknesses, demonstrating the business risks associated with these loopholes and providing risk treatment and prioritization strategies to effectively address the cyber-related risks, threats and vulnerabilities, ensuring appropriate levels of protection, confidentiality, integrity and privacy in alignment with the security framework. General Programming - Applies a computer language to communicate with computers using a set of instructions and to automate the execution of tasks. Intrusion Detection - The use of security analytics, including the outputs from intelligence analysis, predictive research and root cause analysis in order to search for and detect potential breaches or identify recognized indicators and warnings. Also, monitoring and collating external vulnerability reports for organizational relevance, ensuring that relevant vulnerabilities are rectified through formal change processes. Penetration Testing - The practice of testing a computer system, network or web application to find security vulnerabilities that an attacker could exploit. Penetration testing can be automated with software applications or performed manually. Relationship Management - Relationship Management is the conscious aim to develop and manage long-term and/or trusting relationships with internal or external customers, distributors, suppliers, or other parties in an environment which can include marketing, selling, servicing and other areas where a relationship is crucial to on-going success. At a senior level, it includes C-level relationships with senior management. Security Controls - Manages and maintains an information system that focuses on the management of risk and the management of information systems security. Security Governance - The process of developing and disseminating corporate security policies, frameworks and guidelines to ensure that day-to-day business operations are guarded and well protected against risks, threats and vulnerabilities. Security Information & Event Management (SIEM) - A set of tools and services offering real-time visibility across an organization's information security systems, and event log management that consolidates data from numerous sources. Security Policy Management - The process of identifying, implementing, and managing the rules and procedures that all individuals must follow when accessing and using an organization's IT assets and resources. Threat Analysis - Monitor intelligence-gathering and anticipate potential threats to an IT/OT systems proactively. This involves the pre-emptive analysis of potential perpetrators, anomalous activities and evidence-based knowledge and inferences on perpetrators' motivations and tactics. Threat Hunting - Searches through networks, endpoints, and datasets to detect and isolate cyber threats that evade existing security solutions. Vulnerability Management - The process of defining, identifying, classifying and prioritizing vulnerabilities in computer systems, applications and network infrastructures and providing the organization with the necessary knowledge, awareness and risk background to understand the threats to its business MINIMUM QUALIFICATIONS: Bachelor's Degree in Information Technology, related field or equivalent experience. Professional certification, e.g. Security+, Network+, OSCP, GIAC, CEH preferred. 5+ years of relevant experience required As an energy industry leader, our career opportunities fuel personal and professional growth. Location: San Antonio, Texas Additional locations: Findlay, Ohio, Houston, Texas Job Requisition ID: Location Address: 19100 Ridgewood Pkwy Education: Employee Group: Full time Employee Subgroup: . click apply for full job details
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $197,300 - $225,100 for AI Engineer 4 Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $197,300 - $225,100 for AI Engineer 4 Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (AI Foundations, VLM Customization) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations, VLM Customization) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $197,300 - $225,100 for AI Engineer 4 Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (MLX, Agentic AI, Gen AI platform Services) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $197,300 - $225,100 for AI Engineer 4 Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations: LLM Customization, Finetuning, Reinforcement Learning) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (Gen AI Platform Services: Agentic AI, Agent Guardrails, Agent Evaluation, Agent Memory) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Francisco, CA: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (AI Foundations, VLM Customization) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations, VLM Customization) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (AI Foundations, LLM Core and Agentic AI) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations, LLM Core and Agentic AI) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
AI Engineer 4 (AI Foundations, LLM Core and Agentic AI) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/19/2026
Full time
AI Engineer 4 (AI Foundations, LLM Core and Agentic AI) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. What You'll Do: Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products that change how our associates work and how our customers interact with Capital One. Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability, etc. Leverage a broad stack of Open Source and SaaS AI technologies such as AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and more. Invent and introduce state-of-the-art foundation model optimization techniques to improve the performance - scalability, cost, latency, throughput - of large scale production AI systems. Contribute to the technical vision and the long term roadmap of foundational AI systems at Capital One. Own the end-to-end architecture for complex AI systems - ensuring maintainability, observability, and ethical alignment Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are met Mentor Principal and Senior Associates on scalable design, performance tuning and research-to-production translation The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enables you to see and exploit optimization opportunities that others miss You are a resilient trailblazer who can forge new paths to achieve business goals when the route is unknown Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 2 years of experience developing AI and ML algorithms or technologies At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java Preferred Qualifications: Experience leading development AI systems with tradeoff decisions around cost, latency, throughput and accuracy 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, delivering, and supporting AI services Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Experience in building agentic AI systems and agentic workflows Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $197,300 - $225,100 for AI Engineer 4 McLean, VA: $197,300 - $225,100 for AI Engineer 4 New York, NY: $215,200 - $245,600 for AI Engineer 4 San Jose, CA: $215,200 - $245,600 for AI Engineer 4 Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).