Job description: We are seeking a Staff Software Engineer to help develop customer-facing command-and-control software for mission-critical systems. This is a hands-on, full-stack engineering role for someone who thinks in systems, not just features. The successful candidate will lead the architecture and development of production-grade single-page applications and the backend services that support them, including real-time telemetry, operational data, and complex workflows. Responsibilities Architect and implement core systems for complex, data-driven single-page applications. Design backend services, APIs, and data flows supporting real-time telemetry and operational workflows. Lead technical design for full-stack features, ensuring scalable and maintainable implementations. Establish engineering patterns, architectural standards, and code-quality practices. Make foundational decisions on application structure, service boundaries, and system integrations. Resolve performance, reliability, and maintainability issues across the stack. Conduct architectural and code reviews. Mentor engineers through clear technical discussions and constructive feedback. Qualifications Bachelors degree in software engineering, computer science, or a related field, or equivalent relevant experience. 8+ years of engineering experience with a bachelors degree, or 6+ years with a masters degree. Demonstrated experience designing and delivering significant production software systems. Strong full-stack capability, including frontend architecture and backend services. Ability to explain technical tradeoffs and architectural decisions clearly. Experience solving complex engineering problems in a disciplined, high-accountability environment. Ability to write software independently. Familiarity with AI-enabled development tools is welcome but not a substitute for engineering judgment. Preferred Experience Python TypeScript Complex data visualization Design systems Operational or decision-support tools Real-time or mission-critical software environments Work Environment This role is primarily office-based. Strong preference is given to candidates available for onsite collaboration in Sioux Falls, South Dakota. Hybrid arrangements may be considered. Remote candidates will not be prioritized. The role may occasionally require work near production areas with noise, fumes, moving machinery, and varying temperatures. Reasonable accommodations are available for qualified individuals with disabilities. Compensation & Benefits A competitive compensation and benefits package is offered, including health and disability insurance, 401(k) match, flexible spending accounts and HSAs, employee assistance program, tuition reimbursement, parental leave, paid time off, and company-paid holidays. Qualifications: Qualifications Bachelors degree in software engineering, computer science, or a related field, or equivalent relevant experience. 8+ years of engineering experience with a bachelors degree, or 6+ years with a masters degree. Demonstrated experience designing and delivering significant production software systems. Strong full-stack capability, including frontend architecture and backend services. Ability to explain technical tradeoffs and architectural decisions clearly. Experience solving complex engineering problems in a disciplined, high-accountability environment. Ability to write software independently. Familiarity with AI-enabled development tools is welcome but not a substitute for engineering judgment. Preferred Experience Python TypeScript Complex data visualization Design systems Operational or decision-support tools Real-time or mission-critical software environments Work Environment This role is primarily office-based. Strong preference is given to candidates available for onsite collaboration in Sioux Falls, South Dakota. Hybrid arrangements may be considered. Remote candidates will not be prioritized. The role may occasionally require work near production areas with noise, fumes, moving machinery, and varying temperatures. Reasonable accommodations are available for qualified individuals with disabilities. Compensation & Benefits A competitive compensation and benefits package is offered, including health and disability insurance, 401(k) match, flexible spending accounts and HSAs, employee assistance program, tuition reimbursement, parental leave, paid time off, and company-paid holidays. Why is This a Great Opportunity: We are seeking a Staff Software Engineer to help develop customer-facing command-and-control software for mission-critical systems. This is a hands-on, full-stack engineering role for someone who thinks in systems, not just features. The successful candidate will lead the architecture and development of production-grade single-page applications and the backend services that support them, including real-time telemetry, operational data, and complex workflows.
09/22/2026
Full time
Job description: We are seeking a Staff Software Engineer to help develop customer-facing command-and-control software for mission-critical systems. This is a hands-on, full-stack engineering role for someone who thinks in systems, not just features. The successful candidate will lead the architecture and development of production-grade single-page applications and the backend services that support them, including real-time telemetry, operational data, and complex workflows. Responsibilities Architect and implement core systems for complex, data-driven single-page applications. Design backend services, APIs, and data flows supporting real-time telemetry and operational workflows. Lead technical design for full-stack features, ensuring scalable and maintainable implementations. Establish engineering patterns, architectural standards, and code-quality practices. Make foundational decisions on application structure, service boundaries, and system integrations. Resolve performance, reliability, and maintainability issues across the stack. Conduct architectural and code reviews. Mentor engineers through clear technical discussions and constructive feedback. Qualifications Bachelors degree in software engineering, computer science, or a related field, or equivalent relevant experience. 8+ years of engineering experience with a bachelors degree, or 6+ years with a masters degree. Demonstrated experience designing and delivering significant production software systems. Strong full-stack capability, including frontend architecture and backend services. Ability to explain technical tradeoffs and architectural decisions clearly. Experience solving complex engineering problems in a disciplined, high-accountability environment. Ability to write software independently. Familiarity with AI-enabled development tools is welcome but not a substitute for engineering judgment. Preferred Experience Python TypeScript Complex data visualization Design systems Operational or decision-support tools Real-time or mission-critical software environments Work Environment This role is primarily office-based. Strong preference is given to candidates available for onsite collaboration in Sioux Falls, South Dakota. Hybrid arrangements may be considered. Remote candidates will not be prioritized. The role may occasionally require work near production areas with noise, fumes, moving machinery, and varying temperatures. Reasonable accommodations are available for qualified individuals with disabilities. Compensation & Benefits A competitive compensation and benefits package is offered, including health and disability insurance, 401(k) match, flexible spending accounts and HSAs, employee assistance program, tuition reimbursement, parental leave, paid time off, and company-paid holidays. Qualifications: Qualifications Bachelors degree in software engineering, computer science, or a related field, or equivalent relevant experience. 8+ years of engineering experience with a bachelors degree, or 6+ years with a masters degree. Demonstrated experience designing and delivering significant production software systems. Strong full-stack capability, including frontend architecture and backend services. Ability to explain technical tradeoffs and architectural decisions clearly. Experience solving complex engineering problems in a disciplined, high-accountability environment. Ability to write software independently. Familiarity with AI-enabled development tools is welcome but not a substitute for engineering judgment. Preferred Experience Python TypeScript Complex data visualization Design systems Operational or decision-support tools Real-time or mission-critical software environments Work Environment This role is primarily office-based. Strong preference is given to candidates available for onsite collaboration in Sioux Falls, South Dakota. Hybrid arrangements may be considered. Remote candidates will not be prioritized. The role may occasionally require work near production areas with noise, fumes, moving machinery, and varying temperatures. Reasonable accommodations are available for qualified individuals with disabilities. Compensation & Benefits A competitive compensation and benefits package is offered, including health and disability insurance, 401(k) match, flexible spending accounts and HSAs, employee assistance program, tuition reimbursement, parental leave, paid time off, and company-paid holidays. Why is This a Great Opportunity: We are seeking a Staff Software Engineer to help develop customer-facing command-and-control software for mission-critical systems. This is a hands-on, full-stack engineering role for someone who thinks in systems, not just features. The successful candidate will lead the architecture and development of production-grade single-page applications and the backend services that support them, including real-time telemetry, operational data, and complex workflows.
P-57 At Databricks, we are obsessed with enabling data teams to solve the world's toughest problems. We do this by building and running the world's best data and AI infrastructure platform, so our customers can focus on the high value challenges that are central to their own missions. Founded in 2013 by the original creators of Apache Spark, Databricks has grown from a tiny corner office in Berkeley, California to a global organization with over 1000 employees. Thousands of organizations, from small to Fortune 100, trust Databricks with their mission-critical workloads, making us one of the fastest growing SaaS companies in the world. Our engineering teams build highly technical products that fulfill real, important needs in the world. We constantly push the boundaries of data and AI technology, while simultaneously operating with the resilience, security and scale that is critical to making customers successful on our platform. We develop and operate one of the largest scale software platforms. The fleet consists of millions of virtual machines, generating terabytes of logs and processing exabytes of data per day. At our scale, we regularly observe cloud hardware, network, and operating system faults, and our software must gracefully shield our customers from any of the above. As a Data Scientist on the Data Team, you will help build a data-driven culture within Databricks by helping work on top priorities for the company. The Data team also functions as an in-house, production "customer" that dog foods Databricks and drives the future direction of the products. The impact you will have: " Analysis at the speed of thought ": Inform decision making by building robust data science tooling for business leaders, analysts, and other data scientists. " Extend capabilities of Databricks ": Work closely with Data Platform and Product Engineering teams to integrate data science tooling with existing Data team offerings and the core product " Strategic business insights ": Lead insight generation for top company priorities, and key Engineering initiatives (reliability, and efficiency). Gather changing requirements, define project OKRs and milestones, and communicate progress and results to both technical and non-technical audiences. Mentor and guide junior data scientists on the team by helping with project planning, technical decisions, and code and document review. Represent the data science discipline throughout the organization, having a powerful voice to make us more data-driven. Represent Databricks at academic and industrial conferences & events. What we look for: 7+ years of data science, machine learning, advanced analytics experience in high velocity, high-growth companies Extensive experience in applying Data Science / ML in production to build data-driven products for solving business problems. Experience collaborating with and understanding the needs of Senior level stakeholders from a variety of functions including: Engineering, Product, and Technical Operations. Ability to deal with ambiguity in fast paced environments by clarifying requirements and having a keen sense of 0 to 1 solutions. Adept at operating both as an individual contributor and identifying how to orchestrate the build through peers and investments in scalable tooling. Strong coding skills in Python and SQL Experience with distributed data processing systems like Spark and familiarity with software engineering principles around testing, code reviews and deployment. M.S. or Ph.D. in quantitative fields (e.g., Statistics, Math, Computer Science, Physics, Economics, Operational Research or Engineering) Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here. Local Pay Range $192,000-$260,000 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram. Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here. Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.
09/22/2026
Full time
P-57 At Databricks, we are obsessed with enabling data teams to solve the world's toughest problems. We do this by building and running the world's best data and AI infrastructure platform, so our customers can focus on the high value challenges that are central to their own missions. Founded in 2013 by the original creators of Apache Spark, Databricks has grown from a tiny corner office in Berkeley, California to a global organization with over 1000 employees. Thousands of organizations, from small to Fortune 100, trust Databricks with their mission-critical workloads, making us one of the fastest growing SaaS companies in the world. Our engineering teams build highly technical products that fulfill real, important needs in the world. We constantly push the boundaries of data and AI technology, while simultaneously operating with the resilience, security and scale that is critical to making customers successful on our platform. We develop and operate one of the largest scale software platforms. The fleet consists of millions of virtual machines, generating terabytes of logs and processing exabytes of data per day. At our scale, we regularly observe cloud hardware, network, and operating system faults, and our software must gracefully shield our customers from any of the above. As a Data Scientist on the Data Team, you will help build a data-driven culture within Databricks by helping work on top priorities for the company. The Data team also functions as an in-house, production "customer" that dog foods Databricks and drives the future direction of the products. The impact you will have: " Analysis at the speed of thought ": Inform decision making by building robust data science tooling for business leaders, analysts, and other data scientists. " Extend capabilities of Databricks ": Work closely with Data Platform and Product Engineering teams to integrate data science tooling with existing Data team offerings and the core product " Strategic business insights ": Lead insight generation for top company priorities, and key Engineering initiatives (reliability, and efficiency). Gather changing requirements, define project OKRs and milestones, and communicate progress and results to both technical and non-technical audiences. Mentor and guide junior data scientists on the team by helping with project planning, technical decisions, and code and document review. Represent the data science discipline throughout the organization, having a powerful voice to make us more data-driven. Represent Databricks at academic and industrial conferences & events. What we look for: 7+ years of data science, machine learning, advanced analytics experience in high velocity, high-growth companies Extensive experience in applying Data Science / ML in production to build data-driven products for solving business problems. Experience collaborating with and understanding the needs of Senior level stakeholders from a variety of functions including: Engineering, Product, and Technical Operations. Ability to deal with ambiguity in fast paced environments by clarifying requirements and having a keen sense of 0 to 1 solutions. Adept at operating both as an individual contributor and identifying how to orchestrate the build through peers and investments in scalable tooling. Strong coding skills in Python and SQL Experience with distributed data processing systems like Spark and familiarity with software engineering principles around testing, code reviews and deployment. M.S. or Ph.D. in quantitative fields (e.g., Statistics, Math, Computer Science, Physics, Economics, Operational Research or Engineering) Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here. Local Pay Range $192,000-$260,000 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram. Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here. Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.
P-57 At Databricks, we are obsessed with enabling data teams to solve the world's toughest problems, from security threat detection to cancer drug development. We do this by building and running the world's best Data Intelligence Platform, so our customers can focus on the high value challenges that are central to their own missions. Founded in 2013 by the original creators of Apache Spark, Databricks has grown from a tiny corner office in Berkeley, California to a global organization with over 1000 employees. Thousands of organizations, from small to Fortune 100, trust Databricks with their mission-critical workloads, making us one of the fastest growing SaaS companies in the world. Our engineering teams build highly technical products that fulfill real, important needs in the world. We constantly push the boundaries of data and AI technology, while simultaneously operating with the resilience, security and scale that is critical to making customers successful on our platform. We develop and operate one of the largest scale software platforms. The fleet consists of millions of virtual machines, generating terabytes of logs and processing exabytes of data per day. At our scale, we regularly observe cloud hardware, network, and operating system faults, and our software must gracefully shield our customers from any of the above. As a Data Scientist on the Data Team, you will help build a data-driven culture within Databricks by helping solve product and business challenges. The Data team also functions as an in-house, production "customer" that dogfoods Databricks and drives the future direction of the products. The impact you will have: Shape the direction of some of our key data science areas - segmentation, recommendation systems, forecasting, product analytics, churn prediction and insights. Work closely with Engineering, Product Management, Sales and Customer Success to understand product usage patterns and trends and make data-driven decisions, recommendations and forecasts. Manage stakeholders for their focus area - gather changing requirements, define project OKRs and milestones, and communicate progress and results to a non-technical audience. Mentor and guide junior data scientists on the team by helping with project planning, technical decisions, and code and document review. Represent the data science discipline throughout the organization, having a powerful voice to make us more data-driven Build self-serving internal data products to make data simple within the company. Represent Databricks at academic and industrial conferences & events. What we look for: 7+ years of data science, machine learning, advanced analytics experience in high velocity, high-growth companies Extensive experience in applying Data Science / ML for the end-to-end development and deployment of data-driven products for solving business problems. Familiarity with product data science - understanding and tracking customer and user behavior using lenses like adoption, churn, cohorts, segmentation and funnel analysis. Experience collaborating with and understanding the needs of stakeholders from a variety of business functions. We work most closely with Product, Sales and Engineering at the moment, but also work with the Marketing and Finance organizations. Strong coding skills in general purpose languages like Scala or Python, and familiarity with software engineering principles around testing, code reviews and deployment. Proficient in data analysis and visualization using tools like R and Python. Experience with distributed data processing systems like Spark, and proficiency in SQL. MS or Ph.D. in quantitative fields (e.g., Statistics, Math, Computer Science, Physics, Economics, Operational Research or Engineering) Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here. Local Pay Range $192,000-$260,000 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram. Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here. Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.
09/22/2026
Full time
P-57 At Databricks, we are obsessed with enabling data teams to solve the world's toughest problems, from security threat detection to cancer drug development. We do this by building and running the world's best Data Intelligence Platform, so our customers can focus on the high value challenges that are central to their own missions. Founded in 2013 by the original creators of Apache Spark, Databricks has grown from a tiny corner office in Berkeley, California to a global organization with over 1000 employees. Thousands of organizations, from small to Fortune 100, trust Databricks with their mission-critical workloads, making us one of the fastest growing SaaS companies in the world. Our engineering teams build highly technical products that fulfill real, important needs in the world. We constantly push the boundaries of data and AI technology, while simultaneously operating with the resilience, security and scale that is critical to making customers successful on our platform. We develop and operate one of the largest scale software platforms. The fleet consists of millions of virtual machines, generating terabytes of logs and processing exabytes of data per day. At our scale, we regularly observe cloud hardware, network, and operating system faults, and our software must gracefully shield our customers from any of the above. As a Data Scientist on the Data Team, you will help build a data-driven culture within Databricks by helping solve product and business challenges. The Data team also functions as an in-house, production "customer" that dogfoods Databricks and drives the future direction of the products. The impact you will have: Shape the direction of some of our key data science areas - segmentation, recommendation systems, forecasting, product analytics, churn prediction and insights. Work closely with Engineering, Product Management, Sales and Customer Success to understand product usage patterns and trends and make data-driven decisions, recommendations and forecasts. Manage stakeholders for their focus area - gather changing requirements, define project OKRs and milestones, and communicate progress and results to a non-technical audience. Mentor and guide junior data scientists on the team by helping with project planning, technical decisions, and code and document review. Represent the data science discipline throughout the organization, having a powerful voice to make us more data-driven Build self-serving internal data products to make data simple within the company. Represent Databricks at academic and industrial conferences & events. What we look for: 7+ years of data science, machine learning, advanced analytics experience in high velocity, high-growth companies Extensive experience in applying Data Science / ML for the end-to-end development and deployment of data-driven products for solving business problems. Familiarity with product data science - understanding and tracking customer and user behavior using lenses like adoption, churn, cohorts, segmentation and funnel analysis. Experience collaborating with and understanding the needs of stakeholders from a variety of business functions. We work most closely with Product, Sales and Engineering at the moment, but also work with the Marketing and Finance organizations. Strong coding skills in general purpose languages like Scala or Python, and familiarity with software engineering principles around testing, code reviews and deployment. Proficient in data analysis and visualization using tools like R and Python. Experience with distributed data processing systems like Spark, and proficiency in SQL. MS or Ph.D. in quantitative fields (e.g., Statistics, Math, Computer Science, Physics, Economics, Operational Research or Engineering) Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here. Local Pay Range $192,000-$260,000 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram. Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here. Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.
RDQ426R282 Databricks is building the world's best and most secure platform for data and AI. We innovate and deploy industry-leading solutions in security, compliance, and governance. As a member of the Trust and Safety Data Science team, you will work on projects critical to ensuring the security and compliance of the Databricks Platform. Our customers depend on Databricks to keep their data safe, all while orchestrating millions of virtual machines across three clouds in dozens of regions around the globe. Our engineering teams build highly technical products that fulfill real, important needs in the world. We always push the boundaries of data and AI technology, while simultaneously operating with the security and scale that is critical to making customers successful on our platform. We serve many companies with varying security and compliance needs. To efficiently serve these markets, we need to understand how customers use our existing features. This requires data-driven analysis of all aspects of security programs at Databricks. Customers also trust Databricks with their most valuable data and we have the mission to build the most trusted data analytics and ML platform in the world. We're looking to expand our Trust and Safety Data Science team. You will join a group of "full stack" data scientists who partner with engineering and security teams, focusing on strategic plans that make Databricks secure and safe for our customers. The team will use statistical and machine learning techniques for fraud and abuse detection on our platforms using state of the art methods . You can read more about some of our efforts in this blog post. The work in fraud and abuse detection is dynamic and essential, offering an opportunity to make a substantial impact in maintaining the security and efficiency of business operations. More information is available at The impact you will have: You will develop and implement Machine Learning models to detect anomalous activity in products that we offer. You will analyze the performance and pricing of security-related features and work with product and engineering teams to identify important opportunities. You will collaborate with security engineers, trust and safety experts, and machine learning engineers to build a variety of systems and tools that protect Databricks and our customers from threats. You will create solutions and frameworks to meet compliance requirements at Databricks You will gather requirements, define project OKRs and milestones, and communicate progress to both technical and non-technical audiences. You will guide junior data scientists and interns on the team by helping with project planning, technical decisions, and code and document review. You will represent the data science discipline throughout the organization, using your powerful voice to make us more data-driven. You will represent Databricks at academic and industrial conferences and events. What we look for: 7+ years of data science, machine learning, and advanced analytics experience in high-velocity, high-growth companies Understanding of good software engineering practices around testing, code reviews, and deployment. Experience working in a highly cross functional alignment and talking about results to non-technical partners. Experience deploying Data Science / ML solutions in production to achieve results. Coding skills in SQL and a software development language (preferably Python) Experience with distributed data processing systems like Spark and familiarity with software engineering principles. Prior experience applying machine learning and data analytics to identify SaaS product misuse and enhance compliance preferred but not required. Masters or higher in quantitative fields or equivalent experience in industry Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here. Local Pay Range $192,000-$260,000 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram. Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here. Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.
09/22/2026
Full time
RDQ426R282 Databricks is building the world's best and most secure platform for data and AI. We innovate and deploy industry-leading solutions in security, compliance, and governance. As a member of the Trust and Safety Data Science team, you will work on projects critical to ensuring the security and compliance of the Databricks Platform. Our customers depend on Databricks to keep their data safe, all while orchestrating millions of virtual machines across three clouds in dozens of regions around the globe. Our engineering teams build highly technical products that fulfill real, important needs in the world. We always push the boundaries of data and AI technology, while simultaneously operating with the security and scale that is critical to making customers successful on our platform. We serve many companies with varying security and compliance needs. To efficiently serve these markets, we need to understand how customers use our existing features. This requires data-driven analysis of all aspects of security programs at Databricks. Customers also trust Databricks with their most valuable data and we have the mission to build the most trusted data analytics and ML platform in the world. We're looking to expand our Trust and Safety Data Science team. You will join a group of "full stack" data scientists who partner with engineering and security teams, focusing on strategic plans that make Databricks secure and safe for our customers. The team will use statistical and machine learning techniques for fraud and abuse detection on our platforms using state of the art methods . You can read more about some of our efforts in this blog post. The work in fraud and abuse detection is dynamic and essential, offering an opportunity to make a substantial impact in maintaining the security and efficiency of business operations. More information is available at The impact you will have: You will develop and implement Machine Learning models to detect anomalous activity in products that we offer. You will analyze the performance and pricing of security-related features and work with product and engineering teams to identify important opportunities. You will collaborate with security engineers, trust and safety experts, and machine learning engineers to build a variety of systems and tools that protect Databricks and our customers from threats. You will create solutions and frameworks to meet compliance requirements at Databricks You will gather requirements, define project OKRs and milestones, and communicate progress to both technical and non-technical audiences. You will guide junior data scientists and interns on the team by helping with project planning, technical decisions, and code and document review. You will represent the data science discipline throughout the organization, using your powerful voice to make us more data-driven. You will represent Databricks at academic and industrial conferences and events. What we look for: 7+ years of data science, machine learning, and advanced analytics experience in high-velocity, high-growth companies Understanding of good software engineering practices around testing, code reviews, and deployment. Experience working in a highly cross functional alignment and talking about results to non-technical partners. Experience deploying Data Science / ML solutions in production to achieve results. Coding skills in SQL and a software development language (preferably Python) Experience with distributed data processing systems like Spark and familiarity with software engineering principles. Prior experience applying machine learning and data analytics to identify SaaS product misuse and enhance compliance preferred but not required. Masters or higher in quantitative fields or equivalent experience in industry Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here. Local Pay Range $192,000-$260,000 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram. Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here. Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.
Job Description Job Description Lead / Principal Process ML Engineer Job Title: Lead / Principal Process ML Engineer Job Type: Full-time Location: San Diego, CA Summary We are building AI platforms that run on real factory floors, predicting, optimizing, and controlling manufacturing processes in real time. This role is needed to expand our process optimization capability - we need a technical leader who can own the full lifecycle of ML-driven process optimization products, from causal modeling to productization and customer deployment. Currently there is no dedicated technical lead for this product area, and growing customer demand requires a senior leader to set the technical vision and drive execution. Roles and Responsibilities Own the technical direction for key projects - set vision, define what to build, manage priorities, and be accountable for delivery Lead development of ML-driven process optimization products - from causal modeling to productization and customer deployment Make architecture decisions: select modeling approaches, define system boundaries, and evaluate trade-offs between physical fidelity, inference speed, and product constraints Define modeling strategy - decide which physics to encode, which architectures to use, and how to validate against real process data Drive architecture reviews and technical decision-making across the process optimization team Hire, mentor, and grow engineers - build a high-performing team through hiring, code reviews, and technical coaching Coordinate across HQ (Seoul) and overseas R&D labs - aligning research with product roadmaps across time zones Own production-grade delivery - models must run reliably inside equipment operating 24/7 on customer lines Requirements Doctorate (Ph.D.) Mechanical Engineering, Physics, Computer Science, Electrical Engineering, or a related field 8+ years of industry experience post-Ph.D. Proven track record of shipping process optimization or control products from problem definition through customer-facing deployment Experience leading a technical team - setting direction, managing delivery, and making architecture decisions Strong physics foundation (fluid dynamics, thermodynamics, heat transfer, mechanics) Strong ML skills (PyTorch, custom architectures, PINNs, neural operators, surrogate models) Ability to bridge physics and product - translate process understanding into model design and product features Comfort with ambiguity: able to define the problem and the approach, not just execute a spec Experience shipping ML-driven process optimization or control products in an industry setting Experience leading or managing a technical team (ML or applied science) Experience with production software systems - ML integration, data pipelines, deployment infrastructure Preferred Requirements Prior role as Tech Lead, Staff Engineer, or Team Lead in an ML or applied science team Experience productizing process optimization as a commercial product deployed on customer sites Hands-on manufacturing process experience - semiconductor, SMT, electronics assembly, or precision manufacturing Experience coordinating technical direction across distributed teams (HQ + overseas R&D) Published work in scientific machine learning, computational physics, or neural operators Experience with SPC, Cpk analysis, or 3D inspection/metrology systems Benefits • Health/Dental/Vision/Life Insurance at no employee premium (including dependent coverage) • 401K retirement plan with 5% matching • Generous PTO and paid holidays
09/22/2026
Full time
Job Description Job Description Lead / Principal Process ML Engineer Job Title: Lead / Principal Process ML Engineer Job Type: Full-time Location: San Diego, CA Summary We are building AI platforms that run on real factory floors, predicting, optimizing, and controlling manufacturing processes in real time. This role is needed to expand our process optimization capability - we need a technical leader who can own the full lifecycle of ML-driven process optimization products, from causal modeling to productization and customer deployment. Currently there is no dedicated technical lead for this product area, and growing customer demand requires a senior leader to set the technical vision and drive execution. Roles and Responsibilities Own the technical direction for key projects - set vision, define what to build, manage priorities, and be accountable for delivery Lead development of ML-driven process optimization products - from causal modeling to productization and customer deployment Make architecture decisions: select modeling approaches, define system boundaries, and evaluate trade-offs between physical fidelity, inference speed, and product constraints Define modeling strategy - decide which physics to encode, which architectures to use, and how to validate against real process data Drive architecture reviews and technical decision-making across the process optimization team Hire, mentor, and grow engineers - build a high-performing team through hiring, code reviews, and technical coaching Coordinate across HQ (Seoul) and overseas R&D labs - aligning research with product roadmaps across time zones Own production-grade delivery - models must run reliably inside equipment operating 24/7 on customer lines Requirements Doctorate (Ph.D.) Mechanical Engineering, Physics, Computer Science, Electrical Engineering, or a related field 8+ years of industry experience post-Ph.D. Proven track record of shipping process optimization or control products from problem definition through customer-facing deployment Experience leading a technical team - setting direction, managing delivery, and making architecture decisions Strong physics foundation (fluid dynamics, thermodynamics, heat transfer, mechanics) Strong ML skills (PyTorch, custom architectures, PINNs, neural operators, surrogate models) Ability to bridge physics and product - translate process understanding into model design and product features Comfort with ambiguity: able to define the problem and the approach, not just execute a spec Experience shipping ML-driven process optimization or control products in an industry setting Experience leading or managing a technical team (ML or applied science) Experience with production software systems - ML integration, data pipelines, deployment infrastructure Preferred Requirements Prior role as Tech Lead, Staff Engineer, or Team Lead in an ML or applied science team Experience productizing process optimization as a commercial product deployed on customer sites Hands-on manufacturing process experience - semiconductor, SMT, electronics assembly, or precision manufacturing Experience coordinating technical direction across distributed teams (HQ + overseas R&D) Published work in scientific machine learning, computational physics, or neural operators Experience with SPC, Cpk analysis, or 3D inspection/metrology systems Benefits • Health/Dental/Vision/Life Insurance at no employee premium (including dependent coverage) • 401K retirement plan with 5% matching • Generous PTO and paid holidays
Senior Staff Engineer, AI Compute (Remote Eligible) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Capital One machine learning platform organization manages our cloud-based enterprise AI+ML system delivering the high-scale developer and runtime environments required to build, orchestrate, and deploy compute and data intensive AI systems across real-time and batch workloads. We are seeking a Senior Distinguished Engineer, a hands-on technical leader passionate about distributed systems, to engineer and scale foundational compute capabilities for our platform. You will use your experience in building large scale, highly available and high performance systems to develop our common compute infrastructure on top of CPU and GPU substrates. Your contributions will power everything from developer notebooks to ML / DL model training, model inference and feature generation pipelines to pre-training and fine tuning Transformer-based models as well as generative AI inference and agentic applications. Your depth of expertise in technologies including Golang and Python programming languages, popular distributed compute frameworks including Spark / Dask / Ray / Flink, container (e.g., Kubernetes) and serverless (e.g., AWS Lambda) runtime environments, and ML+AI workload patterns will provide an amplifying technical element that is paramount to our team's success. In this role, you will : Architect and build control and data plane implementations required to realize a highly available, multi-tenant, large scale and a secure machine learning platform Develop Ray and Spark distributed compute engine solutions to accelerate diverse workloads from LLM pre-training and reinforcement learning to large-scale data processing, while maximizing compute unit economics Engineer systemic improvements for operational excellence including automating KTLO (Keep The Lights On) workflows Direct the technical execution of a diverse project portfolio, collaborating with developers specializing in everything ranging from distributed microservices to running large foundation models Work cross-functionally with product and program management disciplines, and stakeholder and partners across Capital One to help optimize business outcomes while driving towards strong technology solutions Share your passion for staying on top of tech trends, experimenting with and learning new technologies, participating in internal & external technology communities, and leading system design and code review sessions Help elevate the Capital One Distinguished Engineering community and establish yourself as a go-to resource on given technologies and technology-enabled capabilities Lead the way in creating next-generation talent, mentoring internal talent and actively recruiting external talent to bolster the Capital One tech talent pool Capital One is open to hiring a Remote Employee for this opportunity Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 10 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 8 years of experience developing AI and ML algorithms or technologies At least 10 years of experience programming with Python, Go, Scala, or Java Preferred Qualifications : Master's Degree in Computer Science or a Master's Degree in Software Engineering Hands on experience in the internals of Ray (Actors/GCS/Scheduling) or Spark (Query Optimizer/Memory Management) Experience building platforms that support LLM training, fine-tuning, or high-throughput inference Hands-on experience with AWS-specific compute primitives (EKS, EC2 UltraClusters, Graviton) and cost-optimization strategies History of upstream contributions to major distributed systems projects Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Remote (Regardless of Location): $286,200 - $326,700 for Sr. Distinguished AI Engineer Cambridge, MA: $314,800 - $359,300 for Sr. Distinguished AI Engineer McLean, VA: $314,800 - $359,300 for Sr. Distinguished AI Engineer New York, NY: $343,400 - $392,000 for Sr. Distinguished AI Engineer Richmond, VA: $286,200 - $326,700 for Sr. Distinguished AI Engineer San Francisco, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer San Jose, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
09/21/2026
Full time
Senior Staff Engineer, AI Compute (Remote Eligible) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Capital One machine learning platform organization manages our cloud-based enterprise AI+ML system delivering the high-scale developer and runtime environments required to build, orchestrate, and deploy compute and data intensive AI systems across real-time and batch workloads. We are seeking a Senior Distinguished Engineer, a hands-on technical leader passionate about distributed systems, to engineer and scale foundational compute capabilities for our platform. You will use your experience in building large scale, highly available and high performance systems to develop our common compute infrastructure on top of CPU and GPU substrates. Your contributions will power everything from developer notebooks to ML / DL model training, model inference and feature generation pipelines to pre-training and fine tuning Transformer-based models as well as generative AI inference and agentic applications. Your depth of expertise in technologies including Golang and Python programming languages, popular distributed compute frameworks including Spark / Dask / Ray / Flink, container (e.g., Kubernetes) and serverless (e.g., AWS Lambda) runtime environments, and ML+AI workload patterns will provide an amplifying technical element that is paramount to our team's success. In this role, you will : Architect and build control and data plane implementations required to realize a highly available, multi-tenant, large scale and a secure machine learning platform Develop Ray and Spark distributed compute engine solutions to accelerate diverse workloads from LLM pre-training and reinforcement learning to large-scale data processing, while maximizing compute unit economics Engineer systemic improvements for operational excellence including automating KTLO (Keep The Lights On) workflows Direct the technical execution of a diverse project portfolio, collaborating with developers specializing in everything ranging from distributed microservices to running large foundation models Work cross-functionally with product and program management disciplines, and stakeholder and partners across Capital One to help optimize business outcomes while driving towards strong technology solutions Share your passion for staying on top of tech trends, experimenting with and learning new technologies, participating in internal & external technology communities, and leading system design and code review sessions Help elevate the Capital One Distinguished Engineering community and establish yourself as a go-to resource on given technologies and technology-enabled capabilities Lead the way in creating next-generation talent, mentoring internal talent and actively recruiting external talent to bolster the Capital One tech talent pool Capital One is open to hiring a Remote Employee for this opportunity Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 10 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 8 years of experience developing AI and ML algorithms or technologies At least 10 years of experience programming with Python, Go, Scala, or Java Preferred Qualifications : Master's Degree in Computer Science or a Master's Degree in Software Engineering Hands on experience in the internals of Ray (Actors/GCS/Scheduling) or Spark (Query Optimizer/Memory Management) Experience building platforms that support LLM training, fine-tuning, or high-throughput inference Hands-on experience with AWS-specific compute primitives (EKS, EC2 UltraClusters, Graviton) and cost-optimization strategies History of upstream contributions to major distributed systems projects Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Remote (Regardless of Location): $286,200 - $326,700 for Sr. Distinguished AI Engineer Cambridge, MA: $314,800 - $359,300 for Sr. Distinguished AI Engineer McLean, VA: $314,800 - $359,300 for Sr. Distinguished AI Engineer New York, NY: $343,400 - $392,000 for Sr. Distinguished AI Engineer Richmond, VA: $286,200 - $326,700 for Sr. Distinguished AI Engineer San Francisco, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer San Jose, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
RELOCATION ASSISTANCE: Relocation assistance may be available CLEARANCE REQUIRED FOR START: Yes CLEARANCE TYPE: Secret TRAVEL: Yes, 10% of the Time Description At Northrop Grumman, our employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for generations to come. Our pioneering and inventive spirit has enabled us to be at the forefront of many technological advancements in our nation's history - from the first flight across the Atlantic Ocean, to stealth bombers, to landing on the moon. We look for people who have bold new ideas, courage and a pioneering spirit to join forces to invent the future, and have fun along the way. Our culture thrives on intellectual curiosity, cognitive diversity and bringing your whole self to work - and we have an insatiable drive to do what others think is impossible. Our employees are not only part of history, they're making history. Join Northrop Grumman on our continued mission to push the boundaries of possible across land, sea, air, space, and cyberspace. Enjoy a culture where your voice is valued and start contributing to our team of passionate professionals providing real-life solutions to our world's biggest challenges. We take pride in creating purposeful work and allowing our employees to grow and achieve their goals every day by Defining Possible. With our competitive pay and comprehensive benefits, we have the right opportunities to fit your life and launch your career today. Northrop Grumman Defense Systems is seeking a System States and Controls Capability Lead (Staff Systems Engineer - Level 5) to join the team located in Roy UT, Bellevue NE, Huntsville, AL or Manhattan Beach, CA in support of the Sentinel program. Northrop Grumman supports the Air Force's sustainment, development, production and deployment of hardware and system modifications for Intercontinental Ballistic Missile (ICBM,) Ground and Airborne Launch Control Systems, Launch Facilities, and associated infrastructure. What you will get to do: The Sentinel program has an exciting opportunity for a System States & Controls Capability Lead Staff Systems Engineer (Level 5) to join the Command Systems team leading activities including requirements definition / allocation, functional decomposition, interface definition, verification & validation, and requirements traceability across the ground segment of the Sentinel Weapon System. Specific duties to include, but are not limited to the following: Lead a small team of systems engineers to develop systems engineering artifacts across wing-wide maintenance, operations, initialize, fault detection, and shutdown capabilities Work with both technical teams and stakeholders to develop and mature off-nominal system use cases and states deriving full-scope functionality for the wing and preliminary product selection and development Help develop and incorporate the appropriate requirements for the system and ensure that they are properly represented in the model. Develop systems architecture using Cameo Enterprise Architecture (behavioral, structural, analytical). Contribute to system requirements development, management, and analysis. Develop unifying model techniques, procedures, and processes (for model development, tool integration, and team integration). Basic Qualifications: Bachelor's degree in STEM (Science, Technology, Engineering, and Mathematics) with 12 years of experience; or master's degree with 10 years of experience; or PhD with 8 years of experience Must be a US Citizen with an active DoD Secret Clearance, at time of application, current and within scope, with an investigation date within the last 6 years. Must have the ability to obtain Special Access Program (SAP) approval within a reasonable period, as determined by the company to meet its business needs. 4 years of experience with requirements management/analysis/allocations 3 years of experience with Model-Based Engineering / Model-Based Systems Engineering (MBE/MBSE) practices, languages and/or tools such as Cameo, Rhapsody or MagicDraw 5 years of experience with one or more of the following: Command & Control (C2), physical security, cybersecurity, Nuclear Command, Control, and Communications (NC3) systems/processes, communications systems, facility design, military aerospace development (e.g. Sentinel, Minute Man III, B-2, B-21 delivery systems) Preferred Qualifications: Active DoD Top Secret Clearance Demonstrated understanding of the Systems Engineering Vee Model 3+ years of experience in model-based systems engineering; thread-based system decomposition, requirements engineering, system modeling in SysML Experience in documenting Interface Control Documents, Interface Requirement Specifications, and Interface Description Documents Understanding of Object-Oriented Systems Engineering Methodology and systems thinking Excellent collaboration skills and experience working across product, design, and systems engineering teams with various stakeholder communities Proven technical leadership in designing, modeling, and verifying complex, hierarchical State Machines for distributed, real-time command and control architectures. This includes defining clear, deterministic state transitions for critical mission sequences (e.g., Power-On, Daily Monitoring, Targeting/Planning, Countdown, Launch, Abort, and Shutdown) Extensive experience developing automated FDIR frameworks and Built-In Test (BIT) strategies (Periodic, Initiated, and Continuous) for high-consequence systems. This includes leveraging fault-tree analysis to design algorithms that isolate anomalous behavior down to the specific LRU or software. Demonstrated understanding of the systematic challenges of distributed system synchronization - specifically how a network of geographically separated ground nodes initially coordinates clocks, shares state telemetry, loads cryptographic keys, and handles simultaneous, secure shutdowns Experience with ICBM weapon system engineering and integration Experience with complex system development on large programs As a full-time employee of Northrop Grumman, you are eligible for our robust benefits package including: - Medical, Dental & Vision coverage - 401k - Educational Assistance - Life Insurance - Employee Assistance Programs & Work/Life Solutions - Paid Time Off - Health & Wellness Resources - Employee Discounts This position's standard work schedule is 9/80. The 9/80 schedule allows employees who work a nine-hour day Monday through Thursday to take every other Friday off. Primary Level Salary Range: $152,900.00 - $229,300.00 The above salary range represents a general guideline; however, Northrop Grumman considers a number of factors when determining base salary offers such as the scope and responsibilities of the position and the candidate's experience, education, skills and current market conditions. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay. Annual bonuses are designed to reward individual contributions as well as allow employees to share in company results. Employees in Vice President or Director positions may be eligible for Long Term Incentives. In addition, Northrop Grumman provides a variety of benefits including health insurance coverage, life and disability insurance, savings plan, Company paid holidays and paid time off (PTO) for vacation and/or personal business. The application period for the job is estimated to be 20 days from the job posting date. However, this timeline may be shortened or extended depending on business needs and the availability of qualified candidates. Northrop Grumman is an Equal Opportunity Employer, making decisions without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class. For our complete EEO and pay transparency statement, please visit U.S. Citizenship is required for all positions with a government clearance and certain other restricted positions.
09/20/2026
Full time
RELOCATION ASSISTANCE: Relocation assistance may be available CLEARANCE REQUIRED FOR START: Yes CLEARANCE TYPE: Secret TRAVEL: Yes, 10% of the Time Description At Northrop Grumman, our employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for generations to come. Our pioneering and inventive spirit has enabled us to be at the forefront of many technological advancements in our nation's history - from the first flight across the Atlantic Ocean, to stealth bombers, to landing on the moon. We look for people who have bold new ideas, courage and a pioneering spirit to join forces to invent the future, and have fun along the way. Our culture thrives on intellectual curiosity, cognitive diversity and bringing your whole self to work - and we have an insatiable drive to do what others think is impossible. Our employees are not only part of history, they're making history. Join Northrop Grumman on our continued mission to push the boundaries of possible across land, sea, air, space, and cyberspace. Enjoy a culture where your voice is valued and start contributing to our team of passionate professionals providing real-life solutions to our world's biggest challenges. We take pride in creating purposeful work and allowing our employees to grow and achieve their goals every day by Defining Possible. With our competitive pay and comprehensive benefits, we have the right opportunities to fit your life and launch your career today. Northrop Grumman Defense Systems is seeking a System States and Controls Capability Lead (Staff Systems Engineer - Level 5) to join the team located in Roy UT, Bellevue NE, Huntsville, AL or Manhattan Beach, CA in support of the Sentinel program. Northrop Grumman supports the Air Force's sustainment, development, production and deployment of hardware and system modifications for Intercontinental Ballistic Missile (ICBM,) Ground and Airborne Launch Control Systems, Launch Facilities, and associated infrastructure. What you will get to do: The Sentinel program has an exciting opportunity for a System States & Controls Capability Lead Staff Systems Engineer (Level 5) to join the Command Systems team leading activities including requirements definition / allocation, functional decomposition, interface definition, verification & validation, and requirements traceability across the ground segment of the Sentinel Weapon System. Specific duties to include, but are not limited to the following: Lead a small team of systems engineers to develop systems engineering artifacts across wing-wide maintenance, operations, initialize, fault detection, and shutdown capabilities Work with both technical teams and stakeholders to develop and mature off-nominal system use cases and states deriving full-scope functionality for the wing and preliminary product selection and development Help develop and incorporate the appropriate requirements for the system and ensure that they are properly represented in the model. Develop systems architecture using Cameo Enterprise Architecture (behavioral, structural, analytical). Contribute to system requirements development, management, and analysis. Develop unifying model techniques, procedures, and processes (for model development, tool integration, and team integration). Basic Qualifications: Bachelor's degree in STEM (Science, Technology, Engineering, and Mathematics) with 12 years of experience; or master's degree with 10 years of experience; or PhD with 8 years of experience Must be a US Citizen with an active DoD Secret Clearance, at time of application, current and within scope, with an investigation date within the last 6 years. Must have the ability to obtain Special Access Program (SAP) approval within a reasonable period, as determined by the company to meet its business needs. 4 years of experience with requirements management/analysis/allocations 3 years of experience with Model-Based Engineering / Model-Based Systems Engineering (MBE/MBSE) practices, languages and/or tools such as Cameo, Rhapsody or MagicDraw 5 years of experience with one or more of the following: Command & Control (C2), physical security, cybersecurity, Nuclear Command, Control, and Communications (NC3) systems/processes, communications systems, facility design, military aerospace development (e.g. Sentinel, Minute Man III, B-2, B-21 delivery systems) Preferred Qualifications: Active DoD Top Secret Clearance Demonstrated understanding of the Systems Engineering Vee Model 3+ years of experience in model-based systems engineering; thread-based system decomposition, requirements engineering, system modeling in SysML Experience in documenting Interface Control Documents, Interface Requirement Specifications, and Interface Description Documents Understanding of Object-Oriented Systems Engineering Methodology and systems thinking Excellent collaboration skills and experience working across product, design, and systems engineering teams with various stakeholder communities Proven technical leadership in designing, modeling, and verifying complex, hierarchical State Machines for distributed, real-time command and control architectures. This includes defining clear, deterministic state transitions for critical mission sequences (e.g., Power-On, Daily Monitoring, Targeting/Planning, Countdown, Launch, Abort, and Shutdown) Extensive experience developing automated FDIR frameworks and Built-In Test (BIT) strategies (Periodic, Initiated, and Continuous) for high-consequence systems. This includes leveraging fault-tree analysis to design algorithms that isolate anomalous behavior down to the specific LRU or software. Demonstrated understanding of the systematic challenges of distributed system synchronization - specifically how a network of geographically separated ground nodes initially coordinates clocks, shares state telemetry, loads cryptographic keys, and handles simultaneous, secure shutdowns Experience with ICBM weapon system engineering and integration Experience with complex system development on large programs As a full-time employee of Northrop Grumman, you are eligible for our robust benefits package including: - Medical, Dental & Vision coverage - 401k - Educational Assistance - Life Insurance - Employee Assistance Programs & Work/Life Solutions - Paid Time Off - Health & Wellness Resources - Employee Discounts This position's standard work schedule is 9/80. The 9/80 schedule allows employees who work a nine-hour day Monday through Thursday to take every other Friday off. Primary Level Salary Range: $152,900.00 - $229,300.00 The above salary range represents a general guideline; however, Northrop Grumman considers a number of factors when determining base salary offers such as the scope and responsibilities of the position and the candidate's experience, education, skills and current market conditions. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay. Annual bonuses are designed to reward individual contributions as well as allow employees to share in company results. Employees in Vice President or Director positions may be eligible for Long Term Incentives. In addition, Northrop Grumman provides a variety of benefits including health insurance coverage, life and disability insurance, savings plan, Company paid holidays and paid time off (PTO) for vacation and/or personal business. The application period for the job is estimated to be 20 days from the job posting date. However, this timeline may be shortened or extended depending on business needs and the availability of qualified candidates. Northrop Grumman is an Equal Opportunity Employer, making decisions without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class. For our complete EEO and pay transparency statement, please visit U.S. Citizenship is required for all positions with a government clearance and certain other restricted positions.
Annapurna Labs is an integral part of AWS and develops hardware and software components that are critical building blocks for EC2 infrastructure. We specialize in designing software, systems and chips that optimize the AWS customer experience. The AWS Neuron Collectives team is seeking a Software Engineer to optimize collective operations for AWS Trainium. Trainium is one of Amazon's highest priority initiatives, powering the frontier AI models being trained today. Collectives are the critical operations that scale AI compute across the data center. You'll work in depth to optimize compute for the specific topologies used to train modern LLMs. Working closely with the hardware team, you'll push for maximum performance using C/C++, interfacing with DMA and firmware and investigating detailed topologies. You'll analyze current collective algorithms using publicly accessible tools like Neuron Explorer and optimize these to fully utilize compute and bus bandwidth to scale across the data center. This is a unique opportunity to impact how AI training runs at AWS scale, while growing your technical breadth and depth. Key job responsibilities As a Neuron Collectives Software Developer, you will: Enhance collective algorithms and topologies for optimal training performance Use tools like Neuron Explorer to identify bottlenecks in compute and bus bandwidth utilization Monitor and analyze processor, DMA, firmware, and workload metrics Optimize collective operations to scale AI compute across the data center Work closely with the hardware team to co-optimize software and Trainium silicon Develop and optimize C/C++ implementations of collective communication patterns Investigate and implement improvements for specific training topologies used by modern LLMs Build and maintain analysis frameworks and automation solutions The role offers opportunities to work on cutting-edge AI training hardware while contributing to one of Amazon's most critical initiatives. A day in the life Inclusive Team Culture Here at AWS, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Work/Life Balance Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives. Mentorship & Career Growth Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge sharing and mentorship. We care about your career growth and strive to assign projects based on what will help each team member develop into a better-rounded professional and enable them to take on more complex tasks in the future. About the team Annapurna Labs, part of AWS, created Trainium as a purpose-built AI training chip to revolutionize machine learning at Amazon scale. The Neuron Collectives team owns the software stack that enables collective operations - the communication primitives that allow AI training to scale across thousands of chips in the data center. Our work is essential to training the frontier models that power AI today. We work closely with hardware teams to extract maximum performance from Trainium, ensuring that compute and interconnect bandwidth are fully utilized. Our team sits at the intersection of hardware, firmware, and distributed systems. BASIC QUALIFICATIONS - Experience building complex software systems that have been successfully delivered to customers - Experience contributing to the architecture and design (architecture, design patterns, reliability and scaling) of new and current systems - Bachelor's degree in computer science or equivalent - Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations - Experience in development in the last 3 years, or experience in embedded development in C/C++ PREFERRED QUALIFICATIONS - Master's degree in computer science or equivalent - Experience with hardware/software integration and real-time systems - Familiarity with collective communication algorithms (e.g., all-reduce, all-gather) or distributed training frameworks Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
Annapurna Labs is an integral part of AWS and develops hardware and software components that are critical building blocks for EC2 infrastructure. We specialize in designing software, systems and chips that optimize the AWS customer experience. The AWS Neuron Collectives team is seeking a Software Engineer to optimize collective operations for AWS Trainium. Trainium is one of Amazon's highest priority initiatives, powering the frontier AI models being trained today. Collectives are the critical operations that scale AI compute across the data center. You'll work in depth to optimize compute for the specific topologies used to train modern LLMs. Working closely with the hardware team, you'll push for maximum performance using C/C++, interfacing with DMA and firmware and investigating detailed topologies. You'll analyze current collective algorithms using publicly accessible tools like Neuron Explorer and optimize these to fully utilize compute and bus bandwidth to scale across the data center. This is a unique opportunity to impact how AI training runs at AWS scale, while growing your technical breadth and depth. Key job responsibilities As a Neuron Collectives Software Developer, you will: Enhance collective algorithms and topologies for optimal training performance Use tools like Neuron Explorer to identify bottlenecks in compute and bus bandwidth utilization Monitor and analyze processor, DMA, firmware, and workload metrics Optimize collective operations to scale AI compute across the data center Work closely with the hardware team to co-optimize software and Trainium silicon Develop and optimize C/C++ implementations of collective communication patterns Investigate and implement improvements for specific training topologies used by modern LLMs Build and maintain analysis frameworks and automation solutions The role offers opportunities to work on cutting-edge AI training hardware while contributing to one of Amazon's most critical initiatives. A day in the life Inclusive Team Culture Here at AWS, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Work/Life Balance Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives. Mentorship & Career Growth Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge sharing and mentorship. We care about your career growth and strive to assign projects based on what will help each team member develop into a better-rounded professional and enable them to take on more complex tasks in the future. About the team Annapurna Labs, part of AWS, created Trainium as a purpose-built AI training chip to revolutionize machine learning at Amazon scale. The Neuron Collectives team owns the software stack that enables collective operations - the communication primitives that allow AI training to scale across thousands of chips in the data center. Our work is essential to training the frontier models that power AI today. We work closely with hardware teams to extract maximum performance from Trainium, ensuring that compute and interconnect bandwidth are fully utilized. Our team sits at the intersection of hardware, firmware, and distributed systems. BASIC QUALIFICATIONS - Experience building complex software systems that have been successfully delivered to customers - Experience contributing to the architecture and design (architecture, design patterns, reliability and scaling) of new and current systems - Bachelor's degree in computer science or equivalent - Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations - Experience in development in the last 3 years, or experience in embedded development in C/C++ PREFERRED QUALIFICATIONS - Master's degree in computer science or equivalent - Experience with hardware/software integration and real-time systems - Familiarity with collective communication algorithms (e.g., all-reduce, all-gather) or distributed training frameworks Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the trn and inf servers that use them. This position is for a Software Engineer that will lead the development of various services that will aid in optimization, analysis and release of machine learning workloads and artifacts. This candidate must have had experience leading distributed systems and machine learning related projects, preferably starting from architecture through several generations of delivery to customers. Deep knowledge of optimization, resource management, scheduling are needed. The ideal candidate will have experience working on services like EC2, EKS, Lambda in AWS or similar services on other cloud providers. Key job responsibilities This engineer will lead the design and implementation of new tools, pipelines and automation, will work with developers, system architects, hardware engineers and users both within and external to Amazon to ensure compatibility of this new toolset with existing and next-generation AI accelerators. Design, implement, and maintain CI/CD pipelines to automate the software release process. Collaborate with development teams to integrate new software releases. Infrastructure Management: Manage and automate infrastructure provisioning. Ensure high availability and scalability of systems through effective infrastructure management. Monitoring and Optimization: Implement monitoring solutions to track system performance. Identify bottlenecks and optimize system performance. Security and Compliance: Implement security best practices in the DevOps pipeline. Conduct regular vulnerability assessments and risk management. A day in the life As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also: Build high-impact solutions to deliver to our large customer base. Participate in design discussions, code review, and communicate with internal and external stakeholders. Work cross-functionally to help drive business decisions with your technical input. Work in a startup-like development environment, where you're always working on the most important stuff. About the team The Neuron Infra Services team fosters a builder's culture where experimentation is encouraged, and impact is measurable. We emphasize collaboration, technical ownership, and continuous learning. Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Knowledge of system performance, memory management, and parallel computing principles - Experience in debugging, profiling, and implementing software engineering best practices in large-scale systems, or experience debugging, profiling, and implementing best software engineering practices in large-scale systems - Experience with AWS or cloud technologies PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - Knowledge of fundamentals of networking, security, databases (relational or NoSQL), operating systems (Unix, Linux, and/or Windows) - Fundamentals of Machine learning and LLMs, their architecture along with work experience on certain LLM models. Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the trn and inf servers that use them. This position is for a Software Engineer that will lead the development of various services that will aid in optimization, analysis and release of machine learning workloads and artifacts. This candidate must have had experience leading distributed systems and machine learning related projects, preferably starting from architecture through several generations of delivery to customers. Deep knowledge of optimization, resource management, scheduling are needed. The ideal candidate will have experience working on services like EC2, EKS, Lambda in AWS or similar services on other cloud providers. Key job responsibilities This engineer will lead the design and implementation of new tools, pipelines and automation, will work with developers, system architects, hardware engineers and users both within and external to Amazon to ensure compatibility of this new toolset with existing and next-generation AI accelerators. Design, implement, and maintain CI/CD pipelines to automate the software release process. Collaborate with development teams to integrate new software releases. Infrastructure Management: Manage and automate infrastructure provisioning. Ensure high availability and scalability of systems through effective infrastructure management. Monitoring and Optimization: Implement monitoring solutions to track system performance. Identify bottlenecks and optimize system performance. Security and Compliance: Implement security best practices in the DevOps pipeline. Conduct regular vulnerability assessments and risk management. A day in the life As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also: Build high-impact solutions to deliver to our large customer base. Participate in design discussions, code review, and communicate with internal and external stakeholders. Work cross-functionally to help drive business decisions with your technical input. Work in a startup-like development environment, where you're always working on the most important stuff. About the team The Neuron Infra Services team fosters a builder's culture where experimentation is encouraged, and impact is measurable. We emphasize collaboration, technical ownership, and continuous learning. Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Knowledge of system performance, memory management, and parallel computing principles - Experience in debugging, profiling, and implementing software engineering best practices in large-scale systems, or experience debugging, profiling, and implementing best software engineering practices in large-scale systems - Experience with AWS or cloud technologies PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - Knowledge of fundamentals of networking, security, databases (relational or NoSQL), operating systems (Unix, Linux, and/or Windows) - Fundamentals of Machine learning and LLMs, their architecture along with work experience on certain LLM models. Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The Acceleration Kernel Library team is at the forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance. As part of the broader Neuron Compiler organization, our team works across multiple technology layers - from frameworks and compilers to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of machine learning, high-performance computing, and distributed architectures, where you'll help shape the future of AI acceleration technology This is an opportunity to work on cutting-edge products at the intersection of machine-learning, high-performance computing, and distributed architectures. You will architect and implement business-critical features, publish cutting-edge research, and mentor a brilliant team of experienced engineers. We operate in spaces that are very large, yet our teams remain small and agile. There is no blueprint. We're inventing. We're experimenting. It is a very unique learning culture. The team works closely with customers on their model enablement, providing direct support and optimization expertise to ensure their machine learning workloads achieve optimal performance on AWS ML accelerators. Explore the product and our history! Key job responsibilities Our kernel engineers collaborate across compiler, runtime, framework, and hardware teams to optimize machine learning workloads for our global customer base. Working at the intersection of software, hardware, and machine learning systems, you'll bring expertise in low-level optimization, system architecture, and ML model acceleration. In this role, you will: Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models Analyze and optimize kernel-level performance across multiple generations of Neuron hardware Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks Implement compiler optimizations such as fusion, sharding, tiling, and scheduling Work directly with customers to enable and optimize their ML models on AWS accelerators Collaborate across teams to develop innovative kernel optimization techniques A day in the life As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also: Build high-impact solutions to deliver to our large customer base. Participate in design discussions, code review, and communicate with internal and external stakeholders. Work cross-functionally to help drive business decisions with your technical input. Work in a startup-like development environment, where you're always working on the most important stuff. About the team Amazon Diverse Experiences values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Inclusive Team Culture Here at Amazon, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Work/Life Balance Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The Acceleration Kernel Library team is at the forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance. As part of the broader Neuron Compiler organization, our team works across multiple technology layers - from frameworks and compilers to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of machine learning, high-performance computing, and distributed architectures, where you'll help shape the future of AI acceleration technology This is an opportunity to work on cutting-edge products at the intersection of machine-learning, high-performance computing, and distributed architectures. You will architect and implement business-critical features, publish cutting-edge research, and mentor a brilliant team of experienced engineers. We operate in spaces that are very large, yet our teams remain small and agile. There is no blueprint. We're inventing. We're experimenting. It is a very unique learning culture. The team works closely with customers on their model enablement, providing direct support and optimization expertise to ensure their machine learning workloads achieve optimal performance on AWS ML accelerators. Explore the product and our history! Key job responsibilities Our kernel engineers collaborate across compiler, runtime, framework, and hardware teams to optimize machine learning workloads for our global customer base. Working at the intersection of software, hardware, and machine learning systems, you'll bring expertise in low-level optimization, system architecture, and ML model acceleration. In this role, you will: Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models Analyze and optimize kernel-level performance across multiple generations of Neuron hardware Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks Implement compiler optimizations such as fusion, sharding, tiling, and scheduling Work directly with customers to enable and optimize their ML models on AWS accelerators Collaborate across teams to develop innovative kernel optimization techniques A day in the life As you design and code solutions to help our team drive efficiencies in software architecture, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also: Build high-impact solutions to deliver to our large customer base. Participate in design discussions, code review, and communicate with internal and external stakeholders. Work cross-functionally to help drive business decisions with your technical input. Work in a startup-like development environment, where you're always working on the most important stuff. About the team Amazon Diverse Experiences values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Inclusive Team Culture Here at Amazon, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Work/Life Balance Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
The MLIL DataPlane team is looking for a Software Development Engineer to own the design and implementation of our inference data plane. We build the software that makes large models run efficiently on custom hardware - spanning model execution, memory management, data movement, and serving integration. Our work covers the full inference path: integrating serving engines with custom hardware, developing high-performance compute kernels, enabling efficient data movement, and driving models from early validation through production. We operate at frontier scale with large distributed models. This is a ground-up effort with rapidly evolving hardware and software. We are looking for an individual contributor who can write and optimize low-level code for custom hardware, validate model architectures end-to-end, build test and profiling infrastructure, and drive performance across the stack. Key job responsibilities - Develop and optimize compute kernels for a custom ML accelerator architecture, targeting production-level performance for large language model inference. - Implement and validate LLM architectures end-to-end - from PyTorch model definition through distributed execution on custom hardware. - Integrate custom accelerator backends into open-source ML serving frameworks (vLLM, PyTorch), including scheduler extensions, memory management, and model parallelism. - Build and maintain test infrastructure for model correctness validation across CPU, GPU, simulator, and hardware targets. - Profile and optimize inference workloads - identify bottlenecks, instrument critical paths, and drive latency and throughput improvements from simulation through hardware bring-up. - Own features end-to-end: from design through implementation, testing, and integration into the broader software stack. - Contribute to CI/CD pipelines that gate model and kernel changes on correctness and performance regressions. BASIC QUALIFICATIONS - Bachelor's degree or equivalent - 4+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Knowledge of computer architecture, operating systems, and parallel computing - Knowledge of Linux fundamentals - Strong proficiency in C/C++ - Experience developing compute kernels for GPUs, DSPs, or custom accelerators - Proven track record of owning and delivering complex software features end-to-end PREFERRED QUALIFICATIONS - Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques - Knowledge of ML frameworks including JAX, PyTorch, vLLM, SGLang, Dynamo, TorchXLA, and TensorRT - Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware - Experience with distributed systems - collective communication, RDMA, or high-speed interconnect programming - Experience with hardware simulation environments and model validation workflows - Demonstrated early adopter of AI-assisted development tools - uses LLMs or code-generation agents as part of daily workflow Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
The MLIL DataPlane team is looking for a Software Development Engineer to own the design and implementation of our inference data plane. We build the software that makes large models run efficiently on custom hardware - spanning model execution, memory management, data movement, and serving integration. Our work covers the full inference path: integrating serving engines with custom hardware, developing high-performance compute kernels, enabling efficient data movement, and driving models from early validation through production. We operate at frontier scale with large distributed models. This is a ground-up effort with rapidly evolving hardware and software. We are looking for an individual contributor who can write and optimize low-level code for custom hardware, validate model architectures end-to-end, build test and profiling infrastructure, and drive performance across the stack. Key job responsibilities - Develop and optimize compute kernels for a custom ML accelerator architecture, targeting production-level performance for large language model inference. - Implement and validate LLM architectures end-to-end - from PyTorch model definition through distributed execution on custom hardware. - Integrate custom accelerator backends into open-source ML serving frameworks (vLLM, PyTorch), including scheduler extensions, memory management, and model parallelism. - Build and maintain test infrastructure for model correctness validation across CPU, GPU, simulator, and hardware targets. - Profile and optimize inference workloads - identify bottlenecks, instrument critical paths, and drive latency and throughput improvements from simulation through hardware bring-up. - Own features end-to-end: from design through implementation, testing, and integration into the broader software stack. - Contribute to CI/CD pipelines that gate model and kernel changes on correctness and performance regressions. BASIC QUALIFICATIONS - Bachelor's degree or equivalent - 4+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Knowledge of computer architecture, operating systems, and parallel computing - Knowledge of Linux fundamentals - Strong proficiency in C/C++ - Experience developing compute kernels for GPUs, DSPs, or custom accelerators - Proven track record of owning and delivering complex software features end-to-end PREFERRED QUALIFICATIONS - Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques - Knowledge of ML frameworks including JAX, PyTorch, vLLM, SGLang, Dynamo, TorchXLA, and TensorRT - Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware - Experience with distributed systems - collective communication, RDMA, or high-speed interconnect programming - Experience with hardware simulation environments and model validation workflows - Demonstrated early adopter of AI-assisted development tools - uses LLMs or code-generation agents as part of daily workflow Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
Annapurna Labs designs silicon and software that accelerates innovation. Customers choose us to create cloud solutions that solve challenges that were unimaginable a short time ago-even yesterday. Our custom chips, accelerators, and software stacks enable us to take on technical challenges that have never been seen before, and deliver results that help our customers change the world. The Machine Learning Server Software Team is looking for candidates interested in writing data-driven software for our Machine Learning servers. We build production software to initialize and monitor the most advanced machine learning acceleration servers in the world. We build systems software that initializes custom accelerator chips and monitors peripherals with device drivers for I2C, SPI, PCie and technologies alike. Our team does not work on machine learning algorithms, but rather on the physical systems (hardware) which execute and accelerate those machine learning algorithms. Data paths, PCIe, SPI, I2C, accelerator inner-workings are our bread and butter. Come join our team! Key job responsibilities - Member of a team responsible for the software associated with server components and integration in to EC2. - Working with the MLA Hardware, Test and Manufacturing teams to create a coordinated software package to enable both qualification as well as rapid deployment of software. - Developing software (C, C++, Python, Lua) which can be maintained, improved upon, documented, tested, and reused. A day in the life The MLA Systems Software team was formed to focus on server software primarily for initialization, monitoring, debug, testing, qualification, and manufacturing. At a high-level our goal is to find ways to help the organization scale though the use of software and automation. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually USA, TX, Austin - 143 400.00 USD annually
09/20/2026
Full time
Annapurna Labs designs silicon and software that accelerates innovation. Customers choose us to create cloud solutions that solve challenges that were unimaginable a short time ago-even yesterday. Our custom chips, accelerators, and software stacks enable us to take on technical challenges that have never been seen before, and deliver results that help our customers change the world. The Machine Learning Server Software Team is looking for candidates interested in writing data-driven software for our Machine Learning servers. We build production software to initialize and monitor the most advanced machine learning acceleration servers in the world. We build systems software that initializes custom accelerator chips and monitors peripherals with device drivers for I2C, SPI, PCie and technologies alike. Our team does not work on machine learning algorithms, but rather on the physical systems (hardware) which execute and accelerate those machine learning algorithms. Data paths, PCIe, SPI, I2C, accelerator inner-workings are our bread and butter. Come join our team! Key job responsibilities - Member of a team responsible for the software associated with server components and integration in to EC2. - Working with the MLA Hardware, Test and Manufacturing teams to create a coordinated software package to enable both qualification as well as rapid deployment of software. - Developing software (C, C++, Python, Lua) which can be maintained, improved upon, documented, tested, and reused. A day in the life The MLA Systems Software team was formed to focus on server software primarily for initialization, monitoring, debug, testing, qualification, and manufacturing. At a high-level our goal is to find ways to help the organization scale though the use of software and automation. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually USA, TX, Austin - 143 400.00 USD annually
Every token a large language model generates depends on data reaching the right accelerator at the right moment. As AI models outgrow any single chip, the network between accelerators becomes the bottleneck that decides how fast - and how affordably - the world's largest models can serve real users. That network layer is what our team builds. We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving into separate prefill and decode pools and moving the model's KV cache between them at the limit of what the hardware allows. Get it right and users get answers in milliseconds; get it wrong and the fastest accelerators in the world sit idle waiting on data. You'll build components of the high-speed transfer path that make that difference, and you'll learn to measure success in how close we run to the theoretical peak of the machine. In this role you will: Build and optimize the low-level data-movement software that transfers KV cache and activations across accelerators, servers, and heterogeneous memory - over AWS's highest-performance network fabric. Profile real workloads, find the true bottleneck, and close the gap between "it works" and "it runs fast" - pushing components toward the hardware's limit. Work across the stack - from network transport up to the inference frameworks - learning from the teams building the chips, runtime, and models. Deliver features that ship to our largest clusters, for our largest customers, serving the largest AI models in production. What we're looking for: Strong C/C++ and a genuine interest in low-level, performance-critical systems - solid command of Linux, memory, and writing fast code. The instinct to ask "how fast could this go?" and the discipline to measure it. Exposure to high-speed networking, HPC interconnects, or GPU/accelerator systems (RDMA, InfiniBand, libfabric, UCX, NCCL, MPI) is a strong plus; embedded-systems experience is welcome. Prior AI/ML experience is not required - if you're a strong systems engineer eager to learn, we'll teach you the ML side. If you like solving genuinely hard problems, working alongside HPC and ML customers, iterating fast, and shipping at a scale few places can offer, come join us. You'll work alongside senior engineers and Principal Engineers who've built this layer from the ground up, with real room to grow your scope and technical depth - on a team at the leading edge of AI/ML infrastructure. About the team: You'd be joining Annapurna Labs, an integral part of AWS. Annapurna designs the hardware and software building blocks behind EC2 - every EC2 instance runs on hardware we designed. We specialize in the chips, systems, and software that optimize the AWS customer experience, and this team sits where AI meets the silicon and the network underneath it. A day in the life Annapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS. We have mixed discipline orgs, you'd be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you'll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Experience with C/C++ PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
Every token a large language model generates depends on data reaching the right accelerator at the right moment. As AI models outgrow any single chip, the network between accelerators becomes the bottleneck that decides how fast - and how affordably - the world's largest models can serve real users. That network layer is what our team builds. We're looking for an engineer to work at the frontier of disaggregated inference: splitting LLM serving into separate prefill and decode pools and moving the model's KV cache between them at the limit of what the hardware allows. Get it right and users get answers in milliseconds; get it wrong and the fastest accelerators in the world sit idle waiting on data. You'll build components of the high-speed transfer path that make that difference, and you'll learn to measure success in how close we run to the theoretical peak of the machine. In this role you will: Build and optimize the low-level data-movement software that transfers KV cache and activations across accelerators, servers, and heterogeneous memory - over AWS's highest-performance network fabric. Profile real workloads, find the true bottleneck, and close the gap between "it works" and "it runs fast" - pushing components toward the hardware's limit. Work across the stack - from network transport up to the inference frameworks - learning from the teams building the chips, runtime, and models. Deliver features that ship to our largest clusters, for our largest customers, serving the largest AI models in production. What we're looking for: Strong C/C++ and a genuine interest in low-level, performance-critical systems - solid command of Linux, memory, and writing fast code. The instinct to ask "how fast could this go?" and the discipline to measure it. Exposure to high-speed networking, HPC interconnects, or GPU/accelerator systems (RDMA, InfiniBand, libfabric, UCX, NCCL, MPI) is a strong plus; embedded-systems experience is welcome. Prior AI/ML experience is not required - if you're a strong systems engineer eager to learn, we'll teach you the ML side. If you like solving genuinely hard problems, working alongside HPC and ML customers, iterating fast, and shipping at a scale few places can offer, come join us. You'll work alongside senior engineers and Principal Engineers who've built this layer from the ground up, with real room to grow your scope and technical depth - on a team at the leading edge of AI/ML infrastructure. About the team: You'd be joining Annapurna Labs, an integral part of AWS. Annapurna designs the hardware and software building blocks behind EC2 - every EC2 instance runs on hardware we designed. We specialize in the chips, systems, and software that optimize the AWS customer experience, and this team sits where AI meets the silicon and the network underneath it. A day in the life Annapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS. We have mixed discipline orgs, you'd be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you'll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language - Experience with C/C++ PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
As a Neuron Collectives Software Developer, you will: Enhance collective algorithms and topologies for optimal training performance Use tools like Neuron Explorer to identify bottlenecks in compute and bus bandwidth utilization Monitor and analyze processor, DMA, firmware, and workload metrics Optimize collective operations to scale AI compute across the data center through low level device driver development Work closely with the hardware team to co-optimize software and Trainium silicon Develop and optimize C/C++ implementations of collective communication patterns Investigate and implement improvements for specific training topologies used by modern LLMs Build and maintain analysis frameworks and automation solutions The role offers opportunities to work on cutting-edge AI training hardware while contributing to one of Amazon's most critical initiatives. A day in the life Annapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS. We have mixed discipline orgs, you'd be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you'll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving. About the team Annapurna Labs, part of AWS, created Trainium as a purpose-built AI training chip to revolutionize machine learning at Amazon scale. The Neuron Collectives team owns the software stack that enables collective operations - the communication primitives that allow AI training to scale across thousands of chips in the data center. Our work is essential to training the frontier models that power AI today. We work closely with hardware teams to extract maximum performance from Trainium, ensuring that compute and interconnect bandwidth are fully utilized. Our team sits at the intersection of hardware, firmware, and distributed systems. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
09/20/2026
Full time
As a Neuron Collectives Software Developer, you will: Enhance collective algorithms and topologies for optimal training performance Use tools like Neuron Explorer to identify bottlenecks in compute and bus bandwidth utilization Monitor and analyze processor, DMA, firmware, and workload metrics Optimize collective operations to scale AI compute across the data center through low level device driver development Work closely with the hardware team to co-optimize software and Trainium silicon Develop and optimize C/C++ implementations of collective communication patterns Investigate and implement improvements for specific training topologies used by modern LLMs Build and maintain analysis frameworks and automation solutions The role offers opportunities to work on cutting-edge AI training hardware while contributing to one of Amazon's most critical initiatives. A day in the life Annapurna Labs, a crucial part of AWS, is responsible for developing hardware and software components for EC2 infrastructure. Our team focuses on building networking solutions that for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS. We have mixed discipline orgs, you'd be working side by side with infrastructure experts, hardware engineers, RTL engineers, scientists & architects. Our workforce spans the globe and is truly international, you'll find yourself working side by side with individuals from numerous countries. We take mentorship seriously, you can both expect senior mentorship and will be expected to mentor new and junior engineers. The pace is fast as we work on the latest advancements of AI/ML, but we take the time to bond as a team and enjoy the successes. We offer flexibility in working hours, and respect WLB as a core org tenet. The team enjoys working with numerous principal-level engineers and closely with directors, career growth opportunities are certainly available. This is a role where you will always be encouraged to keep learning, the AI/ML field is fast moving and constantly evolving. About the team Annapurna Labs, part of AWS, created Trainium as a purpose-built AI training chip to revolutionize machine learning at Amazon scale. The Neuron Collectives team owns the software stack that enables collective operations - the communication primitives that allow AI training to scale across thousands of chips in the data center. Our work is essential to training the frontier models that power AI today. We work closely with hardware teams to extract maximum performance from Trainium, ensuring that compute and interconnect bandwidth are fully utilized. Our team sits at the intersection of hardware, firmware, and distributed systems. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually
Annapurna Labs designs silicon and software that accelerates innovation. Customers choose us to create cloud solutions that solve challenges that were unimaginable a short time ago-even yesterday. Our custom chips, accelerators, and software stacks enable us to take on technical challenges that have never been seen before, and deliver results that help our customers change the world. The Machine Learning Server Software Team is looking for candidates interested in writing data-driven software for our Machine Learning servers. We build production software to initialize and monitor the most advanced machine learning acceleration servers in the world. We touch technologies including accelerator chip initialization, setting up clocks and voltages, to systems level device drivers for I2C infrastructure pervasive in the server and everything in between. Our team does not work on machine learning algorithms, but rather on the physical systems (hardware) which execute and accelerate those machine learning algorithms. Data paths, PCIe, SPI, I2C, accelerator inner-workings are our bread and butter. Come join our team. Key job responsibilities - Member of a team responsible for the software associated with server components and integration in to EC2. - Working with the MLA Hardware, Test and Manufacturing teams to create a coordinated software package to enable both qualification as well as rapid deployment of software. - Developing software (C, C++, Python, Lua) which can be maintained, improved upon, documented, tested, and reused. A day in the life The MLA Systems Software team was formed to focus on server software primarily for initialization, monitoring, debug, testing, qualification, and manufacturing. At a high-level our goal is to find ways to help the organization scale though the use of software and automation. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually USA, TX, Austin - 143 400.00 USD annually
09/20/2026
Full time
Annapurna Labs designs silicon and software that accelerates innovation. Customers choose us to create cloud solutions that solve challenges that were unimaginable a short time ago-even yesterday. Our custom chips, accelerators, and software stacks enable us to take on technical challenges that have never been seen before, and deliver results that help our customers change the world. The Machine Learning Server Software Team is looking for candidates interested in writing data-driven software for our Machine Learning servers. We build production software to initialize and monitor the most advanced machine learning acceleration servers in the world. We touch technologies including accelerator chip initialization, setting up clocks and voltages, to systems level device drivers for I2C infrastructure pervasive in the server and everything in between. Our team does not work on machine learning algorithms, but rather on the physical systems (hardware) which execute and accelerate those machine learning algorithms. Data paths, PCIe, SPI, I2C, accelerator inner-workings are our bread and butter. Come join our team. Key job responsibilities - Member of a team responsible for the software associated with server components and integration in to EC2. - Working with the MLA Hardware, Test and Manufacturing teams to create a coordinated software package to enable both qualification as well as rapid deployment of software. - Developing software (C, C++, Python, Lua) which can be maintained, improved upon, documented, tested, and reused. A day in the life The MLA Systems Software team was formed to focus on server software primarily for initialization, monitoring, debug, testing, qualification, and manufacturing. At a high-level our goal is to find ways to help the organization scale though the use of software and automation. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. About AWS Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Mentorship & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually USA, TX, Austin - 143 400.00 USD annually
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the servers that use them. As the Software Development Engineer for the Neuron Runtime Team, you will be responsible for working alongside a team of engineers to develop and maintain high-performance runtime libraries and drivers for machine learning applications and AI accelerators. You will work on design, development, and deployment of Neuron Runtime and other Neuron components. The profiler plays a crucial role to internal and external customers in optimizing AI workloads across hardware platforms such as Trainium and Inferentia devices, by providing deep insights into performance bottlenecks and system behavior. Improving performance of ML Kernels and ML Frameworks. In this role, you will manage the full development life cycle of the Neuron Runtime, ensuring scalability, reliability, and usability. You will collaborate with cross-functional teams to ensure that the our C++ compiler generates key information so customers can understand and optimize the performance of our custom hardware. Additionally, you will drive innovations that allow the profiler to support multiple frameworks, such as PyTorch, JAX, and XLA. A successful candidate will have experience in architecting, building, and operating distributed systems with a focus on high availability and fault tolerance, Hands-on experience with AWS services (e.g., EC2, ECS, CloudWatch, S3, Lambda) in production environments and track record in Owning services end-to-end including deployment, monitoring, alarming, on-call, and post-incident review. A day in the life You will work with the executive leadership and other senior management and technical leaders to define product directions and deliver them to customers. We build massive-scale distributed training and inference solutions. This organization builds the full stack of software, servers and chips to accelerate at the highest scale. About the team Here at AWS, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Work/Life Balance Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives. Mentorship & Career Growth Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge sharing and mentorship. We care about your career growth and strive to assign projects based on what will help each team member develop into a better-rounded professional and enable them to take on more complex tasks in the future. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually USA, WA, Seattle - 143 400.00 USD annually
09/20/2026
Full time
AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the servers that use them. As the Software Development Engineer for the Neuron Runtime Team, you will be responsible for working alongside a team of engineers to develop and maintain high-performance runtime libraries and drivers for machine learning applications and AI accelerators. You will work on design, development, and deployment of Neuron Runtime and other Neuron components. The profiler plays a crucial role to internal and external customers in optimizing AI workloads across hardware platforms such as Trainium and Inferentia devices, by providing deep insights into performance bottlenecks and system behavior. Improving performance of ML Kernels and ML Frameworks. In this role, you will manage the full development life cycle of the Neuron Runtime, ensuring scalability, reliability, and usability. You will collaborate with cross-functional teams to ensure that the our C++ compiler generates key information so customers can understand and optimize the performance of our custom hardware. Additionally, you will drive innovations that allow the profiler to support multiple frameworks, such as PyTorch, JAX, and XLA. A successful candidate will have experience in architecting, building, and operating distributed systems with a focus on high availability and fault tolerance, Hands-on experience with AWS services (e.g., EC2, ECS, CloudWatch, S3, Lambda) in production environments and track record in Owning services end-to-end including deployment, monitoring, alarming, on-call, and post-incident review. A day in the life You will work with the executive leadership and other senior management and technical leaders to define product directions and deliver them to customers. We build massive-scale distributed training and inference solutions. This organization builds the full stack of software, servers and chips to accelerate at the highest scale. About the team Here at AWS, we embrace our differences. We are committed to furthering our culture of inclusion. We have ten employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We have innovative benefit offerings, and host annual and ongoing learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences. Amazon's culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Work/Life Balance Our team puts a high value on work-life balance. It isn't about how many hours you spend at home or at work; it's about the flow you establish that brings energy to both parts of your life. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We offer flexibility in working hours and encourage you to find your own balance between your work and personal lives. Mentorship & Career Growth Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge sharing and mentorship. We care about your career growth and strive to assign projects based on what will help each team member develop into a better-rounded professional and enable them to take on more complex tasks in the future. BASIC QUALIFICATIONS - 3+ years of non-internship professional software development experience - 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience - Experience programming with at least one software programming language PREFERRED QUALIFICATIONS - 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, CA, Cupertino - 165 600.00 USD annually USA, WA, Seattle - 143 400.00 USD annually
National Radio Astronomy Observatory
Green Bank, West Virginia
National Radio Astronomy Observatory Title: Software Engineer II-III Location: 1011 Lopezville Rd, Socorro, NM 87801, USA• 5651 Balloon Fiesta Pkwy, Albuquerque, NM 87113, USA• 155 Observatory Rd, Green Bank, WV 24944, USA Requisition Number: 279 Job Family: Software Engineer Pay Type: Salary Required Education: CPP Position Description: Position Summary The National Radio Astronomy Observatory (NRAO) is a prestigious research and development organization that plays a vital role in the study of the universe. Associated Universities, Inc. (AUI) is a nonprofit organization that manages and operates the NRAO under a cooperative agreement with the National Science Foundation. The Observatory is a hub for technological and scientific collaboration, operating state-of-the-art radio telescope facilities for use by the international scientific community. The Observatory has been instrumental in the study of black holes, galaxies, and the early universe. NRAO is seeking an experienced Software Engineer to join the Online Software Group. The online system is the real-time heart of our observing software. It is responsible for configuring, controlling, and monitoring the real-time systems of our telescopes, including the the Very Large Array (VLA) in New Mexico, the Very Long Baseline Array (VLBA) spread across the US and its territories, and the Green Bank Telescope (GBT) in West Virginia. This software is responsible for taking scientific-domain configuration parameters and decomposing them into the hardware settings and motion commands required by our antennas, and into the corresponding configurations for the associated digital signal processing instruments (e.g. VEGAS for the GBT, and correlators for the VLA and the VLBA). When an observation runs, this software is what makes it happen. This position will be based in Albuquerque, NM, Socorro, NM, Charlottesville, VA or Green Bank, WV. For well-qualified candidates, a remote work arrangement may be considered. What You Will be Doing Designing, implementing, testing, and maintaining components of the online software that configure and control the VLA, VLBA, and GBT in real time. Develop and extend the translation layer that decomposes scientific-domain observing parameters into the engineering parameters the hardware understands: IF-chain configurations, antenna motion commands, and digital signal processing configurations for instruments like VEGAS, WIDAR, and DiFX. Model the observing domain in software - frequency setups, subband and baseband layouts, polarization products, timing and integration structure - and implement the rules, constraints, and legality checks that determine which configurations are achievable on which instrument. Work from interface control documents and hardware specifications, in close collaboration with the engineers and scientists who own the underlying subsystems, to ensure the commands the online system emits are correct and complete. Diagnose and resolve issues that span the layers between a scientist's observing specification and the command stream sent to the telescope, often in collaboration with telescope operators, engineers, and scientific staff. Support the evolution of the online system to meet new observing paradigms and instrumentation, including work that informs and feeds into the next generation Very Large Array (ngVLA). Write and maintain critical documentation, including requirements, software design documents, interface control documents, and user documentation. Participate in code review, testing, and release processes for software that runs mission-critical, 24/7 scientific facilities. Work Environment The successful candidate will join a team of professionals engaged in research and development in the fields of science, engineering, software development, and education. Work is typically performed in a research or development environment. Must be able to operate a personal computer. Occasional travel (domestic and international) may be required. Who You Are: Education A Bachelor's degree in computer science, engineering, scientific or related field; highly relevant experience may be considered in lieu of a Bachelor's degree. While not required, you may have an advanced degree in a related field. Experience One or more years of experience developing software applications. Candidates with progressively more experience will be considered for a higher-level position ranking. Relevant experience with radio astronomy operating software and procedures is preferred. Knowledge of radio astronomy theories and practice would be valuable. Demonstrated experience providing technical leadership of complex data acquisition, scheduling, and operational support systems is preferred. Skills and Competencies Proficiency with Java; familiarity with C/C++ and Python is valuable A solid understanding of object-oriented design and development Demonstrated ability to learn and apply new software languages and unfamiliar domains Understanding of networking concepts and technologies: multicast, TCP, UDP, HTTP, XML, JSON, REST Experience with version control software, testing methodologies, and CI/CD Strong interpersonal and communication skills, including the ability to work effectively with scientists and hardware engineers Comfort working in a Linux (RHEL) environment The following are highly preferred: Experience translating high-level domain concepts into machine- or hardware-level representations - compilers, planners, schedulers, configuration engines, or similar systems that turn intent into instructions Skill in domain modeling and in expressing complex validation rules and constraints clearly in code Experience with distributed systems, or with control and monitoring software for scientific, industrial, or other physical instruments A working comfort with mathematics and unit-bearing quantities - coordinate transformations, frequency and time calculations, and the care required to get them right Familiarity with signal processing concepts, digital IF chains, or correlator architectures Familiarity with basic astronomical principles, radio astronomy, or interferometry Experience collaborating in an Agile/Scrum environment, including sprint planning, estimating tasks, and breaking down requirements into user stories Note that this position does not involve direct hardware or firmware development; though an understanding of these systems and real-time control principles can be valuable. Your work will live at the layer above, producing correct hardware-level configurations and commands from scientific intent, against interfaces defined and maintained by others. Additional Requirement Observatory employees must be authorized to work in the United States. The Observatory presently cannot sponsor H-1B Visas for this position. Total Rewards: Compensation The starting salary of this position is between $87,000 and $121,000. Factors which may affect starting pay within this range may include; education, experience, skills, competencies, other qualifications of the successful candidate, as well as internal equity and labor market conditions. Benefits: Associated Universities, Inc (AUI) offers a comprehensive benefits package addressing the needs of employees and their families with most benefits beginning on the first day of employment, subject to eligibility requirements. AUI provides: Excellent paid time off (13 holidays, annual accrual of up to 24 vacation days) Medical, dental and vision plans are effective on the first day of employment. AUI's retirement benefit contributes an amount equal to 10 percent of a qualified participant's base pay with no required employee contribution. Click Total Rewards for more information. Application Instructions: Select the "Apply" button above. Please be prepared to upload your current CV/Resume and a cover letter describing interest and suitability for the position . Equal Opportunity Employer Statement: AUI is an equal opportunity employer. To view our complete statement, please visit . If you require reasonable accommodation for any part of the application or hiring process, you may submit your request by sending an email to . PM20 Compensation details: 00 Yearly Salary PIf669a7c2591a-1646
09/19/2026
Full time
National Radio Astronomy Observatory Title: Software Engineer II-III Location: 1011 Lopezville Rd, Socorro, NM 87801, USA• 5651 Balloon Fiesta Pkwy, Albuquerque, NM 87113, USA• 155 Observatory Rd, Green Bank, WV 24944, USA Requisition Number: 279 Job Family: Software Engineer Pay Type: Salary Required Education: CPP Position Description: Position Summary The National Radio Astronomy Observatory (NRAO) is a prestigious research and development organization that plays a vital role in the study of the universe. Associated Universities, Inc. (AUI) is a nonprofit organization that manages and operates the NRAO under a cooperative agreement with the National Science Foundation. The Observatory is a hub for technological and scientific collaboration, operating state-of-the-art radio telescope facilities for use by the international scientific community. The Observatory has been instrumental in the study of black holes, galaxies, and the early universe. NRAO is seeking an experienced Software Engineer to join the Online Software Group. The online system is the real-time heart of our observing software. It is responsible for configuring, controlling, and monitoring the real-time systems of our telescopes, including the the Very Large Array (VLA) in New Mexico, the Very Long Baseline Array (VLBA) spread across the US and its territories, and the Green Bank Telescope (GBT) in West Virginia. This software is responsible for taking scientific-domain configuration parameters and decomposing them into the hardware settings and motion commands required by our antennas, and into the corresponding configurations for the associated digital signal processing instruments (e.g. VEGAS for the GBT, and correlators for the VLA and the VLBA). When an observation runs, this software is what makes it happen. This position will be based in Albuquerque, NM, Socorro, NM, Charlottesville, VA or Green Bank, WV. For well-qualified candidates, a remote work arrangement may be considered. What You Will be Doing Designing, implementing, testing, and maintaining components of the online software that configure and control the VLA, VLBA, and GBT in real time. Develop and extend the translation layer that decomposes scientific-domain observing parameters into the engineering parameters the hardware understands: IF-chain configurations, antenna motion commands, and digital signal processing configurations for instruments like VEGAS, WIDAR, and DiFX. Model the observing domain in software - frequency setups, subband and baseband layouts, polarization products, timing and integration structure - and implement the rules, constraints, and legality checks that determine which configurations are achievable on which instrument. Work from interface control documents and hardware specifications, in close collaboration with the engineers and scientists who own the underlying subsystems, to ensure the commands the online system emits are correct and complete. Diagnose and resolve issues that span the layers between a scientist's observing specification and the command stream sent to the telescope, often in collaboration with telescope operators, engineers, and scientific staff. Support the evolution of the online system to meet new observing paradigms and instrumentation, including work that informs and feeds into the next generation Very Large Array (ngVLA). Write and maintain critical documentation, including requirements, software design documents, interface control documents, and user documentation. Participate in code review, testing, and release processes for software that runs mission-critical, 24/7 scientific facilities. Work Environment The successful candidate will join a team of professionals engaged in research and development in the fields of science, engineering, software development, and education. Work is typically performed in a research or development environment. Must be able to operate a personal computer. Occasional travel (domestic and international) may be required. Who You Are: Education A Bachelor's degree in computer science, engineering, scientific or related field; highly relevant experience may be considered in lieu of a Bachelor's degree. While not required, you may have an advanced degree in a related field. Experience One or more years of experience developing software applications. Candidates with progressively more experience will be considered for a higher-level position ranking. Relevant experience with radio astronomy operating software and procedures is preferred. Knowledge of radio astronomy theories and practice would be valuable. Demonstrated experience providing technical leadership of complex data acquisition, scheduling, and operational support systems is preferred. Skills and Competencies Proficiency with Java; familiarity with C/C++ and Python is valuable A solid understanding of object-oriented design and development Demonstrated ability to learn and apply new software languages and unfamiliar domains Understanding of networking concepts and technologies: multicast, TCP, UDP, HTTP, XML, JSON, REST Experience with version control software, testing methodologies, and CI/CD Strong interpersonal and communication skills, including the ability to work effectively with scientists and hardware engineers Comfort working in a Linux (RHEL) environment The following are highly preferred: Experience translating high-level domain concepts into machine- or hardware-level representations - compilers, planners, schedulers, configuration engines, or similar systems that turn intent into instructions Skill in domain modeling and in expressing complex validation rules and constraints clearly in code Experience with distributed systems, or with control and monitoring software for scientific, industrial, or other physical instruments A working comfort with mathematics and unit-bearing quantities - coordinate transformations, frequency and time calculations, and the care required to get them right Familiarity with signal processing concepts, digital IF chains, or correlator architectures Familiarity with basic astronomical principles, radio astronomy, or interferometry Experience collaborating in an Agile/Scrum environment, including sprint planning, estimating tasks, and breaking down requirements into user stories Note that this position does not involve direct hardware or firmware development; though an understanding of these systems and real-time control principles can be valuable. Your work will live at the layer above, producing correct hardware-level configurations and commands from scientific intent, against interfaces defined and maintained by others. Additional Requirement Observatory employees must be authorized to work in the United States. The Observatory presently cannot sponsor H-1B Visas for this position. Total Rewards: Compensation The starting salary of this position is between $87,000 and $121,000. Factors which may affect starting pay within this range may include; education, experience, skills, competencies, other qualifications of the successful candidate, as well as internal equity and labor market conditions. Benefits: Associated Universities, Inc (AUI) offers a comprehensive benefits package addressing the needs of employees and their families with most benefits beginning on the first day of employment, subject to eligibility requirements. AUI provides: Excellent paid time off (13 holidays, annual accrual of up to 24 vacation days) Medical, dental and vision plans are effective on the first day of employment. AUI's retirement benefit contributes an amount equal to 10 percent of a qualified participant's base pay with no required employee contribution. Click Total Rewards for more information. Application Instructions: Select the "Apply" button above. Please be prepared to upload your current CV/Resume and a cover letter describing interest and suitability for the position . Equal Opportunity Employer Statement: AUI is an equal opportunity employer. To view our complete statement, please visit . If you require reasonable accommodation for any part of the application or hiring process, you may submit your request by sending an email to . PM20 Compensation details: 00 Yearly Salary PIf669a7c2591a-1646
RELOCATION ASSISTANCE: Relocation assistance may be available CLEARANCE REQUIRED FOR START: Yes CLEARANCE TYPE: Secret TRAVEL: Yes, 10% of the Time Description At Northrop Grumman, our employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for generations to come. Our pioneering and inventive spirit has enabled us to be at the forefront of many technological advancements in our nation's history - from the first flight across the Atlantic Ocean, to stealth bombers, to landing on the moon. We look for people who have bold new ideas, courage and a pioneering spirit to join forces to invent the future, and have fun along the way. Our culture thrives on intellectual curiosity, cognitive diversity and bringing your whole self to work - and we have an insatiable drive to do what others think is impossible. Our employees are not only part of history, they're making history. Join Northrop Grumman on our continued mission to push the boundaries of possible across land, sea, air, space, and cyberspace. Enjoy a culture where your voice is valued and start contributing to our team of passionate professionals providing real-life solutions to our world's biggest challenges. We take pride in creating purposeful work and allowing our employees to grow and achieve their goals every day by Defining Possible. With our competitive pay and comprehensive benefits, we have the right opportunities to fit your life and launch your career today. Northrop Grumman Defense Systems is seeking a Facility & Platform Integration Capability Lead (Staff Systems Engineer - Level 5) to join the team located in Roy UT, Bellevue NE, Huntsville, AL or Manhattan Beach, CA in support of the Sentinel program. Northrop Grumman supports the Air Force's sustainment, development, production and deployment of hardware and system modifications for Intercontinental Ballistic Missile (ICBM,) Ground and Airborne Launch Control Systems, Launch Facilities, and associated infrastructure. What you will get to do: The Sentinel program has an exciting opportunity for a Facility & Platform Integration Capability Lead Staff Systems Engineer (Level 5) to join the Command Systems team leading activities including requirements definition / allocation, functional decomposition, interface definition, verification & validation, and requirements traceability across the ground segment of the Sentinel Weapon System. Specific duties to include, but are not limited to the following: Lead a small team of systems engineers to develop systems engineering artifacts across the Electromagnetic Environmental Effects (E3) and Ground Systems threads Work with both technical teams and stakeholders to develop and mature allocation of shielding / filtering / grounding / TEMPEST requirements and associated facility architecture considerations, including grounding schematics and capital electrical design integration across Command and Launch Systems Help develop and incorporate the appropriate requirements for the system and ensure that they are properly represented in the model. Develop systems architecture using Cameo Enterprise Architecture (behavioral, structural, analytical). Contribute to system requirements development, management, and analysis. Develop unifying model techniques, procedures, and processes (for model development, tool integration, and team integration). Basic Qualifications: Bachelor's degree in STEM (Science, Technology, Engineering, and Mathematics) with 12 years of experience; or master's degree with 10 years of experience; or PhD with 8 years of experience Must be a US Citizen with an active DoD Secret Clearance, at time of application, current and within scope, with an investigation date within the last 6 years. Must have the ability to obtain Special Access Program (SAP) approval within a reasonable period, as determined by the company to meet its business needs. 4 years of experience with requirements management/analysis/allocations 3 years of experience with Model-Based Engineering / Model-Based Systems Engineering (MBE/MBSE) practices, languages and/or tools such as Cameo, Rhapsody or MagicDraw 5 years of experience with one or more of the following: Command & Control (C2), physical security, cybersecurity, Nuclear Command, Control, and Communications (NC3) systems/processes, communications systems, facility design, military aerospace development (e.g. Sentinel, Minute Man III, B-2, B-21 delivery systems) Preferred Qualifications: Active DoD Top Secret Clearance Demonstrated understanding of the Systems Engineering Vee Model 3+ years of experience in model-based systems engineering; thread-based system decomposition, requirements engineering, system modeling in SysML Experience in documenting Interface Control Documents, Interface Requirement Specifications, and Interface Description Documents Understanding of Object-Oriented Systems Engineering Methodology and systems thinking Excellent collaboration skills and experience working across product, design, and systems engineering teams with various stakeholder communities Proven technical leadership in designing, modeling, and verifying complex, hierarchical State Machines for distributed, real-time command and control architectures. This includes defining clear, deterministic state transitions for critical mission sequences (e.g., Power-On, Daily Monitoring, Targeting/Planning, Countdown, Launch, Abort, and Shutdown) Extensive experience developing automated FDIR frameworks and Built-In Test (BIT) strategies (Periodic, Initiated, and Continuous) for high-consequence systems. This includes leveraging fault-tree analysis to design algorithms that isolate anomalous behavior down to the specific LRU or software. Demonstrated understanding of the systematic challenges of distributed system synchronization - specifically how a network of geographically separated ground nodes initially coordinates clocks, shares state telemetry, loads cryptographic keys, and handles simultaneous, secure shutdowns Experience with ICBM weapon system engineering and integration Experience with complex system development on large programs As a full-time employee of Northrop Grumman, you are eligible for our robust benefits package including: - Medical, Dental & Vision coverage - 401k - Educational Assistance - Life Insurance - Employee Assistance Programs & Work/Life Solutions - Paid Time Off - Health & Wellness Resources - Employee Discounts This position's standard work schedule is 9/80. The 9/80 schedule allows employees who work a nine-hour day Monday through Thursday to take every other Friday off. Primary Level Salary Range: $152,900.00 - $229,300.00 The above salary range represents a general guideline; however, Northrop Grumman considers a number of factors when determining base salary offers such as the scope and responsibilities of the position and the candidate's experience, education, skills and current market conditions. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay. Annual bonuses are designed to reward individual contributions as well as allow employees to share in company results. Employees in Vice President or Director positions may be eligible for Long Term Incentives. In addition, Northrop Grumman provides a variety of benefits including health insurance coverage, life and disability insurance, savings plan, Company paid holidays and paid time off (PTO) for vacation and/or personal business. The application period for the job is estimated to be 20 days from the job posting date. However, this timeline may be shortened or extended depending on business needs and the availability of qualified candidates. Northrop Grumman is an Equal Opportunity Employer, making decisions without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class. For our complete EEO and pay transparency statement, please visit U.S. Citizenship is required for all positions with a government clearance and certain other restricted positions.
09/18/2026
Full time
RELOCATION ASSISTANCE: Relocation assistance may be available CLEARANCE REQUIRED FOR START: Yes CLEARANCE TYPE: Secret TRAVEL: Yes, 10% of the Time Description At Northrop Grumman, our employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for generations to come. Our pioneering and inventive spirit has enabled us to be at the forefront of many technological advancements in our nation's history - from the first flight across the Atlantic Ocean, to stealth bombers, to landing on the moon. We look for people who have bold new ideas, courage and a pioneering spirit to join forces to invent the future, and have fun along the way. Our culture thrives on intellectual curiosity, cognitive diversity and bringing your whole self to work - and we have an insatiable drive to do what others think is impossible. Our employees are not only part of history, they're making history. Join Northrop Grumman on our continued mission to push the boundaries of possible across land, sea, air, space, and cyberspace. Enjoy a culture where your voice is valued and start contributing to our team of passionate professionals providing real-life solutions to our world's biggest challenges. We take pride in creating purposeful work and allowing our employees to grow and achieve their goals every day by Defining Possible. With our competitive pay and comprehensive benefits, we have the right opportunities to fit your life and launch your career today. Northrop Grumman Defense Systems is seeking a Facility & Platform Integration Capability Lead (Staff Systems Engineer - Level 5) to join the team located in Roy UT, Bellevue NE, Huntsville, AL or Manhattan Beach, CA in support of the Sentinel program. Northrop Grumman supports the Air Force's sustainment, development, production and deployment of hardware and system modifications for Intercontinental Ballistic Missile (ICBM,) Ground and Airborne Launch Control Systems, Launch Facilities, and associated infrastructure. What you will get to do: The Sentinel program has an exciting opportunity for a Facility & Platform Integration Capability Lead Staff Systems Engineer (Level 5) to join the Command Systems team leading activities including requirements definition / allocation, functional decomposition, interface definition, verification & validation, and requirements traceability across the ground segment of the Sentinel Weapon System. Specific duties to include, but are not limited to the following: Lead a small team of systems engineers to develop systems engineering artifacts across the Electromagnetic Environmental Effects (E3) and Ground Systems threads Work with both technical teams and stakeholders to develop and mature allocation of shielding / filtering / grounding / TEMPEST requirements and associated facility architecture considerations, including grounding schematics and capital electrical design integration across Command and Launch Systems Help develop and incorporate the appropriate requirements for the system and ensure that they are properly represented in the model. Develop systems architecture using Cameo Enterprise Architecture (behavioral, structural, analytical). Contribute to system requirements development, management, and analysis. Develop unifying model techniques, procedures, and processes (for model development, tool integration, and team integration). Basic Qualifications: Bachelor's degree in STEM (Science, Technology, Engineering, and Mathematics) with 12 years of experience; or master's degree with 10 years of experience; or PhD with 8 years of experience Must be a US Citizen with an active DoD Secret Clearance, at time of application, current and within scope, with an investigation date within the last 6 years. Must have the ability to obtain Special Access Program (SAP) approval within a reasonable period, as determined by the company to meet its business needs. 4 years of experience with requirements management/analysis/allocations 3 years of experience with Model-Based Engineering / Model-Based Systems Engineering (MBE/MBSE) practices, languages and/or tools such as Cameo, Rhapsody or MagicDraw 5 years of experience with one or more of the following: Command & Control (C2), physical security, cybersecurity, Nuclear Command, Control, and Communications (NC3) systems/processes, communications systems, facility design, military aerospace development (e.g. Sentinel, Minute Man III, B-2, B-21 delivery systems) Preferred Qualifications: Active DoD Top Secret Clearance Demonstrated understanding of the Systems Engineering Vee Model 3+ years of experience in model-based systems engineering; thread-based system decomposition, requirements engineering, system modeling in SysML Experience in documenting Interface Control Documents, Interface Requirement Specifications, and Interface Description Documents Understanding of Object-Oriented Systems Engineering Methodology and systems thinking Excellent collaboration skills and experience working across product, design, and systems engineering teams with various stakeholder communities Proven technical leadership in designing, modeling, and verifying complex, hierarchical State Machines for distributed, real-time command and control architectures. This includes defining clear, deterministic state transitions for critical mission sequences (e.g., Power-On, Daily Monitoring, Targeting/Planning, Countdown, Launch, Abort, and Shutdown) Extensive experience developing automated FDIR frameworks and Built-In Test (BIT) strategies (Periodic, Initiated, and Continuous) for high-consequence systems. This includes leveraging fault-tree analysis to design algorithms that isolate anomalous behavior down to the specific LRU or software. Demonstrated understanding of the systematic challenges of distributed system synchronization - specifically how a network of geographically separated ground nodes initially coordinates clocks, shares state telemetry, loads cryptographic keys, and handles simultaneous, secure shutdowns Experience with ICBM weapon system engineering and integration Experience with complex system development on large programs As a full-time employee of Northrop Grumman, you are eligible for our robust benefits package including: - Medical, Dental & Vision coverage - 401k - Educational Assistance - Life Insurance - Employee Assistance Programs & Work/Life Solutions - Paid Time Off - Health & Wellness Resources - Employee Discounts This position's standard work schedule is 9/80. The 9/80 schedule allows employees who work a nine-hour day Monday through Thursday to take every other Friday off. Primary Level Salary Range: $152,900.00 - $229,300.00 The above salary range represents a general guideline; however, Northrop Grumman considers a number of factors when determining base salary offers such as the scope and responsibilities of the position and the candidate's experience, education, skills and current market conditions. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay. Annual bonuses are designed to reward individual contributions as well as allow employees to share in company results. Employees in Vice President or Director positions may be eligible for Long Term Incentives. In addition, Northrop Grumman provides a variety of benefits including health insurance coverage, life and disability insurance, savings plan, Company paid holidays and paid time off (PTO) for vacation and/or personal business. The application period for the job is estimated to be 20 days from the job posting date. However, this timeline may be shortened or extended depending on business needs and the availability of qualified candidates. Northrop Grumman is an Equal Opportunity Employer, making decisions without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class. For our complete EEO and pay transparency statement, please visit U.S. Citizenship is required for all positions with a government clearance and certain other restricted positions.
Development InfoStructure
Washington, Washington DC
Job Description Job Description Join Our TeamDevis is a leading provider of innovative software development, management, and consulting services, specializing in cutting-edge technologies such as DevSecOps, AI, and Machine Learning. With over 30 years of experience, we have established ourselves as a trusted partner for government agencies, delivering tailored, mission-critical solutions that drive digital transformation and operational excellence. Our client-centric approach, coupled with our deep domain expertise and technical prowess, enables us to forge enduring relationships and consistently deliver high-impact, adaptive solutions that resonate with the unique needs of the public sector.The RoleThis is the second audiovisual and video teleconferencing seat on the helpdesk team supporting a federal agency, working alongside the senior specialist, the production engineer, and a part-time event coordinator to keep conference rooms, executive meetings, and live events running. The role is hands-on: a mix of scheduled event support, walk-up troubleshooting, and preventive maintenance. The agency runs Microsoft Teams, Zoom, Webex, Google Meet, and Slack for video conferencing and collaboration, with Cisco and Poly platforms in its conference and AV rooms. Headquarters is in Washington, DC, with 53 field offices nationwide; about 2,100 people rely on these systems.What You'll Do Configure, operate, and troubleshoot AV equipment across conference rooms and event spaces: mixers, DSPs, control systems, projection, and LED displays Schedule, set up, and manage Teams, Zoom, Webex, and Poly RealPresence sessions, including webinars and breakout rooms, and resolve connectivity and interoperability issues Assist with setup, execution, and post-event work for town halls, executive meetings, training sessions, webcasts, and conferences, operating production equipment under the direction of the production engineer or event coordinator Support recording, streaming, public address, public displays, and captioning during events Work the AV/VTC service desk alongside the senior specialist: receive, document, and resolve trouble tickets from staff at any agency location, escalating what you cannot close Perform preventive maintenance, firmware updates, and inspections, and keep inventory records current Help end users get comfortable with conference room technology What You'll Need Six years supporting enterprise AV and VTC operations at comparable scale (the agency runs about 2,100 users across headquarters and 53 field offices) Hands-on experience operating conference room AV and video teleconferencing systems Experience supporting live meetings and events in a professional or government setting Comfort diagnosing codec, connectivity, and hardware issues under time pressure Nice to Have AVIXA CTS certification Prior support to a federal agency AV-over-IP experience Experience delivering end user training on conference room technology Section 508 accessibility experience, particularly captioning ITIL v4 Foundations and experience with Zendesk or a comparable ticketing platform Location, Schedule, and Logistics Standard schedule is an eight-hour day scheduled between 7:00 a.m. and 7:00 p.m. Eastern, Monday through Friday, excluding federal holidays. Occasional weekend work may come up with advance government approval. Live events sometimes land at the edges of the 7:00 a.m. to 7:00 p.m. window, so expect early starts or late wraps around major events. On-site at agency headquarters in Washington, DC (NoMa area, Metro-accessible). Telework is by exception only and requires government approval. You will work on a government-furnished laptop. Personal devices cannot connect to the agency network. At least 40 hours per year of paid professional training in your core skill areas, at no cost to you. Occasional travel to agency field offices, event sites, and official events held away from headquarters, reimbursed at federal travel rates. Devis is an AA/EOE/M/F/Disabled/VET Employer committed to providing equal employment opportunity without regard to an individual's race, color, religion, age, gender, sexual orientation, veteran status, national origin or disability. Powered by JazzHR C15oYRmRPm
09/17/2026
Full time
Job Description Job Description Join Our TeamDevis is a leading provider of innovative software development, management, and consulting services, specializing in cutting-edge technologies such as DevSecOps, AI, and Machine Learning. With over 30 years of experience, we have established ourselves as a trusted partner for government agencies, delivering tailored, mission-critical solutions that drive digital transformation and operational excellence. Our client-centric approach, coupled with our deep domain expertise and technical prowess, enables us to forge enduring relationships and consistently deliver high-impact, adaptive solutions that resonate with the unique needs of the public sector.The RoleThis is the second audiovisual and video teleconferencing seat on the helpdesk team supporting a federal agency, working alongside the senior specialist, the production engineer, and a part-time event coordinator to keep conference rooms, executive meetings, and live events running. The role is hands-on: a mix of scheduled event support, walk-up troubleshooting, and preventive maintenance. The agency runs Microsoft Teams, Zoom, Webex, Google Meet, and Slack for video conferencing and collaboration, with Cisco and Poly platforms in its conference and AV rooms. Headquarters is in Washington, DC, with 53 field offices nationwide; about 2,100 people rely on these systems.What You'll Do Configure, operate, and troubleshoot AV equipment across conference rooms and event spaces: mixers, DSPs, control systems, projection, and LED displays Schedule, set up, and manage Teams, Zoom, Webex, and Poly RealPresence sessions, including webinars and breakout rooms, and resolve connectivity and interoperability issues Assist with setup, execution, and post-event work for town halls, executive meetings, training sessions, webcasts, and conferences, operating production equipment under the direction of the production engineer or event coordinator Support recording, streaming, public address, public displays, and captioning during events Work the AV/VTC service desk alongside the senior specialist: receive, document, and resolve trouble tickets from staff at any agency location, escalating what you cannot close Perform preventive maintenance, firmware updates, and inspections, and keep inventory records current Help end users get comfortable with conference room technology What You'll Need Six years supporting enterprise AV and VTC operations at comparable scale (the agency runs about 2,100 users across headquarters and 53 field offices) Hands-on experience operating conference room AV and video teleconferencing systems Experience supporting live meetings and events in a professional or government setting Comfort diagnosing codec, connectivity, and hardware issues under time pressure Nice to Have AVIXA CTS certification Prior support to a federal agency AV-over-IP experience Experience delivering end user training on conference room technology Section 508 accessibility experience, particularly captioning ITIL v4 Foundations and experience with Zendesk or a comparable ticketing platform Location, Schedule, and Logistics Standard schedule is an eight-hour day scheduled between 7:00 a.m. and 7:00 p.m. Eastern, Monday through Friday, excluding federal holidays. Occasional weekend work may come up with advance government approval. Live events sometimes land at the edges of the 7:00 a.m. to 7:00 p.m. window, so expect early starts or late wraps around major events. On-site at agency headquarters in Washington, DC (NoMa area, Metro-accessible). Telework is by exception only and requires government approval. You will work on a government-furnished laptop. Personal devices cannot connect to the agency network. At least 40 hours per year of paid professional training in your core skill areas, at no cost to you. Occasional travel to agency field offices, event sites, and official events held away from headquarters, reimbursed at federal travel rates. Devis is an AA/EOE/M/F/Disabled/VET Employer committed to providing equal employment opportunity without regard to an individual's race, color, religion, age, gender, sexual orientation, veteran status, national origin or disability. Powered by JazzHR C15oYRmRPm
Network Services Technician Position Number: B258PD Starting Wage/Salary: $70,000 - $80,000 plus exceptional benefits Close Date: Open Until Filled: Yes Open Until Filled Notes: Priority Application Deadline 9/13/2026. Please note interviews for this position will only be conducted in person at our Bend Campus. Incomplete applications will not be considered by hiring committees. Primary Purpose: Responsible for assisting the Network Services Manager with the day-to-day functioning of the College local and wide-area data/video/voice networks including troubleshooting routers, hubs, switches, and communication lines. This position also supports other ITS teams in evaluating, configuring, and administering the COCC network. The position is part of the ITS Network Team within the Tech Support Services area of the ITS department. As such, will also participate in research, development, and direct support of projects that enhance COCC's information technology. Essential Duties and Responsibilities: Assist with network maintenance and overall network stability by: Performing daily operations and monitoring a wired and wireless network. Assisting the Network Services Manager with the installation, maintenance, and monitoring of networking equipment. This may include applying software updates as well as major upgrades. Identifying trends, recommending and implementing corrective measures, and documenting action taken. Serving as a technical specialist in network problems and emergencies; assists in troubleshooting and resolution of network problems. Assisting with analysis and monitoring of network activities to ensure optimal network operation. Helping install and improve network infrastructure to include fiber, cabling, and equipment. Maintaining network documentation and wiring documentation pertaining to the network systems. Helping maintain an inventory of network equipment and parts. Assisting in the maintenance of phone and voice mail systems. Assisting with cross-team training and support of end-user technical issues including, workstations, AV programming, and hardware/software engineering. Network Infrastructure Support COCC learning environment by: Working with the campus constituents, network team, and ITS team members to ensure stable and functional networking services are provided to all COCC campuses. Identifying, suggesting, and implementing improvements to systems and network environment. Assists the Network Services Manager to provide a stable and reliable network environment between State Universities and COCC Provide technical assistance for existing and new building facilities in the support of future technologies. Coordinate with outside vendors as necessary to obtain technical support, quotes, and guidance for network activities. Performing network maintenance during designated maintenance windows. Network Projects As part of the networking team, evaluate and implement projects aimed at sustaining the following: Crestron programming and connectivity Wireless and Wired Network Configuration, Troubleshooting, and Support. Utilizing Network Monitoring tools Firewall Monitoring Network Documentation Power / UPS Management Phone and Voice Mail Support Open Community Wireless Network Support Residence Hall Network Troubleshooting and Support Outlying Campus Support. Network Support of Outside Services AV Switch and Equipment Programming and Configuration. Network Cabling Installation, Troubleshooting, and Documentation. Knowledge, Skills, and Abilities: Individuals must possess these knowledge, skills, and abilities or be able to explain and demonstrate that the individual can perform the essential functions of the job, with or without reasonable accommodation, using some other combination of skills and abilities. A strong understanding of Network Topology and client connections to the network is required. This individual blends strong interpersonal skills with excellent technical ability. Ability to perform classroom AV installs and programming. Understand the issues involved with administering and maintaining educational infrastructure, including network connectivity, Internet access, wireless access, etc. Ability to provide input in making decisions regarding changes to the network such that interruptions are rare, generally planned, and brief. Ability to review and evaluate the College networks and make recommendations. Ability to gather information from other ITS staff to obtain information regarding potentially related problems to network systems. Basic knowledge of data security including encryption, intrusion detection, firewalls, virus protection, etc. Ability to use office equipment, power tools, machinery, computers, and network diagnostic tools. Ability to work with Network and Communication Service providers and ITS technical support staff. Knowledge monitoring and maintaining routers, switches, and other networking devices. Ability to support outside agencies telecommunications needs as defined by contractional agreements with COCC . Must be able to communicate effectively, both orally and in writing, using the English language with or without the use of an interpreter. Minimum Requirements: Education Associate's degree in a technology-related field or equivalent experience. Experience Three years of experience with network infrastructure activities such as monitoring and configuring network routers and switches, programming room control systems, and supporting network infrastructure. One year of experience working in an enterprise and/or production technology environment, providing complex end-user support. Other Valid Oregon Driver's license, and the ability to meet the College requirements to drive campus vehicles; or the ability to obtain within 30-days of employment. Preferred Qualifications: Education Bachelor's degree in a technology-related field or equivalent. Experience Greater than three years of experience with networking infrastructure activities. Experience with wiring network infrastructure including types of wiring categories. To apply, visit The goal of Central Oregon Community College is to provide an atmosphere that encourages our faculty, staff and students to realize their full potential. In support of this goal, it is the policy of Central Oregon Community College that there will be no discrimination or harassment on the basis of age, disability, sex, marital status, national origin, ethnicity, color, race, religion, sexual orientation, gender identity, genetic information, citizenship status, veteran or military status, pregnancy or any other classes protected under federal and state statutes in any education program, activities or employment. Persons with questions about this statement should contact Human Resources at or the Vice President for Student Affairs at . This policy covers nondiscrimination in both employment and access to educational opportunities. When brought to the attention of the appropriate parties, any such actions will be promptly and equitably responded to according to the process outlined in general procedures sections N-1, N-2, or N-3. In support of COCC's EEO statement, bilingual fluency in English and Spanish is considered a plus, along with experience working in a diverse multicultural setting. Copyright 2025 Inc. All rights reserved. Posted by the FREE value-added recruitment advertising agency je-65b0f4457be0a994b8d648b26
09/17/2026
Full time
Network Services Technician Position Number: B258PD Starting Wage/Salary: $70,000 - $80,000 plus exceptional benefits Close Date: Open Until Filled: Yes Open Until Filled Notes: Priority Application Deadline 9/13/2026. Please note interviews for this position will only be conducted in person at our Bend Campus. Incomplete applications will not be considered by hiring committees. Primary Purpose: Responsible for assisting the Network Services Manager with the day-to-day functioning of the College local and wide-area data/video/voice networks including troubleshooting routers, hubs, switches, and communication lines. This position also supports other ITS teams in evaluating, configuring, and administering the COCC network. The position is part of the ITS Network Team within the Tech Support Services area of the ITS department. As such, will also participate in research, development, and direct support of projects that enhance COCC's information technology. Essential Duties and Responsibilities: Assist with network maintenance and overall network stability by: Performing daily operations and monitoring a wired and wireless network. Assisting the Network Services Manager with the installation, maintenance, and monitoring of networking equipment. This may include applying software updates as well as major upgrades. Identifying trends, recommending and implementing corrective measures, and documenting action taken. Serving as a technical specialist in network problems and emergencies; assists in troubleshooting and resolution of network problems. Assisting with analysis and monitoring of network activities to ensure optimal network operation. Helping install and improve network infrastructure to include fiber, cabling, and equipment. Maintaining network documentation and wiring documentation pertaining to the network systems. Helping maintain an inventory of network equipment and parts. Assisting in the maintenance of phone and voice mail systems. Assisting with cross-team training and support of end-user technical issues including, workstations, AV programming, and hardware/software engineering. Network Infrastructure Support COCC learning environment by: Working with the campus constituents, network team, and ITS team members to ensure stable and functional networking services are provided to all COCC campuses. Identifying, suggesting, and implementing improvements to systems and network environment. Assists the Network Services Manager to provide a stable and reliable network environment between State Universities and COCC Provide technical assistance for existing and new building facilities in the support of future technologies. Coordinate with outside vendors as necessary to obtain technical support, quotes, and guidance for network activities. Performing network maintenance during designated maintenance windows. Network Projects As part of the networking team, evaluate and implement projects aimed at sustaining the following: Crestron programming and connectivity Wireless and Wired Network Configuration, Troubleshooting, and Support. Utilizing Network Monitoring tools Firewall Monitoring Network Documentation Power / UPS Management Phone and Voice Mail Support Open Community Wireless Network Support Residence Hall Network Troubleshooting and Support Outlying Campus Support. Network Support of Outside Services AV Switch and Equipment Programming and Configuration. Network Cabling Installation, Troubleshooting, and Documentation. Knowledge, Skills, and Abilities: Individuals must possess these knowledge, skills, and abilities or be able to explain and demonstrate that the individual can perform the essential functions of the job, with or without reasonable accommodation, using some other combination of skills and abilities. A strong understanding of Network Topology and client connections to the network is required. This individual blends strong interpersonal skills with excellent technical ability. Ability to perform classroom AV installs and programming. Understand the issues involved with administering and maintaining educational infrastructure, including network connectivity, Internet access, wireless access, etc. Ability to provide input in making decisions regarding changes to the network such that interruptions are rare, generally planned, and brief. Ability to review and evaluate the College networks and make recommendations. Ability to gather information from other ITS staff to obtain information regarding potentially related problems to network systems. Basic knowledge of data security including encryption, intrusion detection, firewalls, virus protection, etc. Ability to use office equipment, power tools, machinery, computers, and network diagnostic tools. Ability to work with Network and Communication Service providers and ITS technical support staff. Knowledge monitoring and maintaining routers, switches, and other networking devices. Ability to support outside agencies telecommunications needs as defined by contractional agreements with COCC . Must be able to communicate effectively, both orally and in writing, using the English language with or without the use of an interpreter. Minimum Requirements: Education Associate's degree in a technology-related field or equivalent experience. Experience Three years of experience with network infrastructure activities such as monitoring and configuring network routers and switches, programming room control systems, and supporting network infrastructure. One year of experience working in an enterprise and/or production technology environment, providing complex end-user support. Other Valid Oregon Driver's license, and the ability to meet the College requirements to drive campus vehicles; or the ability to obtain within 30-days of employment. Preferred Qualifications: Education Bachelor's degree in a technology-related field or equivalent. Experience Greater than three years of experience with networking infrastructure activities. Experience with wiring network infrastructure including types of wiring categories. To apply, visit The goal of Central Oregon Community College is to provide an atmosphere that encourages our faculty, staff and students to realize their full potential. In support of this goal, it is the policy of Central Oregon Community College that there will be no discrimination or harassment on the basis of age, disability, sex, marital status, national origin, ethnicity, color, race, religion, sexual orientation, gender identity, genetic information, citizenship status, veteran or military status, pregnancy or any other classes protected under federal and state statutes in any education program, activities or employment. Persons with questions about this statement should contact Human Resources at or the Vice President for Student Affairs at . This policy covers nondiscrimination in both employment and access to educational opportunities. When brought to the attention of the appropriate parties, any such actions will be promptly and equitably responded to according to the process outlined in general procedures sections N-1, N-2, or N-3. In support of COCC's EEO statement, bilingual fluency in English and Spanish is considered a plus, along with experience working in a diverse multicultural setting. Copyright 2025 Inc. All rights reserved. Posted by the FREE value-added recruitment advertising agency je-65b0f4457be0a994b8d648b26