it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

3 jobs found

Email me jobs like this
Refine Search
Current Search
bmc engineer
Operations Engineer, MetalDev
CoreWeave New York, New York
Job Description Job Description CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $109,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
09/29/2026
Full time
Job Description Job Description CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $109,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
Operations Engineer, MetalDev
CoreWeave Livingston, New Jersey
Job Description Job Description CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $109,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
09/29/2026
Full time
Job Description Job Description CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $109,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
Boeing
Lead Systems Engineer - Phantom Works
Boeing Saint Louis, Missouri
Job Description At Boeing, we innovate and collaborate to make the world a better place. We're committed to fostering an environment for every teammate that's welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us. Boeing Defense, Space & Security (BDS) is seeking a Lead Systems Engineer (Level 4) to support the Space Battle Management, Command and Control (SBMC2) programs in Colorado Springs, CO or Berkeley, MO. The Systems Engineer will perform as part of a high-performing team and have the opportunity to contribute to the development and delivery of innovative software solutions in an agile software development environment interfacing with internal and external stakeholders; to include a Joint Industry partnership Team (JIPT) and Working Groups, contributing to the design and development of next generation ground command and control capabilities. As a member of the Boeing team, you will be responsible for key portions of our development lifecycle, from idea creation and development, all the way through to maintenance and support of the customer's delivered system. More importantly, you will have the opportunity to make an impact on the results of our projects. We offer a collaborative mentoring environment where you have the opportunity to learn from others and be a mentor to others. Position Responsibilities Support definition of requirements, interfaces, and concept of operations through supporting and/or chairing working group meetings. Understand and communicate prioritization of efforts to multi-disciplinary team. Think abstractly and see the big picture while evaluating technical details Perform technical analyses to develop and validate models of system behavior; identify solutions to complex problems. Serve as a direct interface to both internal and external customers Conduct and support trade studies. Communicate across multiple functional groups (ex. Software, DevSecOps teams, hardware, cybersecurity, IT/infrastructure, networking, mission engineering, etc) Work in an industry consortium with diverse viewpoints Travel may be required up to 10% of the time; Domestically and/or internationally depending on business needs. This position requires an active U.S. Top Secret Security Clearance (U.S. Citizenship Required. (A U.S. Security Clearance that has been active in the past 24 months is considered active) Basic Qualifications (Required Skills/Experience) Bachelor of Science degree from an accredited course of study in engineering, engineering technology (includes manufacturing engineering technology), chemistry, physics, mathematics, data science, or computer science 9+ years of work-related engineering experience 3+ years of experience working on a software development effort, supporting the development of products from inception through delivery and operation (full product lifecycle) Experience with an interface standard, preferably Open Mission Systems (OMS), Universal Command and Control Interface (UCI) experience or Open Architecture standards Experience with Agile development Preferred Qualifications (Desired Skills/Experience) Satellite operations experience Familiarity with Space Domain Awareness Basic modeling skills with SysML, Enterprise Architect, or UML Knowledge of the space mission domain including launch vehicles, space systems, mission and space vehicle command & control (C2), specialized mission payloads, system integration, and interoperability. Conflict of Interest Successful candidates for this job must satisfy the Company's Conflict of Interest (COI) assessment process. Drug Free Workplace Boeing is a Drug Free Workplace where post offer applicants and employees are subject to testing for marijuana, cocaine, opioids, amphetamines, PCP, and alcohol when criteria is met as outlined in our policies. Employee Referral Referral to this job is eligible for bonus to qualifying candidates. Total Rewards At Boeing, we strive to deliver a Total Rewards package that will attract, engage and retain the top talent. Elements of the Total Rewards package include competitive base pay and variable compensation opportunities. The Boeing Company also provides eligible employees with an opportunity to enroll in a variety of benefit programs, generally including health insurance, flexible spending accounts, health savings accounts, retirement savings plans, life and disability insurance programs, and a number of programs that provide for both paid and unpaid time away from work. The specific programs and options available to any given employee may vary depending on eligibility factors such as geographic location, date of hire, and the applicability of collective bargaining agreements. The Boeing 401(k) helps you save for your future, with contributions from Boeing that can help you grow your retirement savings. Our best-in-class retirement benefit features: Best in class 401(k) plan: we'll match your contributions dollar for dollar, up to 10% of eligible pay with Immediate 100% vesting Student Loan Match: The Boeing 401(k) Student Loan Match allows eligible enrolled U.S. employees to have their qualified student loan debt payments counted, along with any match-eligible contributions they make, for purposes of determining the Company Match to employees' Boeing 401(k) accounts. Pay is based upon candidate experience and qualifications, as well as market and business considerations. Summary pay range: $136,850 - $185,150 Applications for this position will be accepted until Oct. 07, 2026 Export Control Requirements: This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a "U.S. Person" as defined by 22 C.F.R. 120.62 is required. "U.S. Person" includes U.S. Citizen, U.S. National, lawful permanent resident, refugee, or asylee. Export Control Details: US based job, US Person required Education Bachelor's Degree or Equivalent Required Relocation This position offers relocation based on candidate eligibility. Security Clearance This position requires an active U.S. Top Secret Security Clearance (U.S. Citizenship Required). (A U.S. Security Clearance that has been active in the past 24 months is considered active) Visa Sponsorship Employer will not sponsor applicants for employment visa status. Shift This position is for 1st shift Equal Opportunity Employer: Boeing is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, national origin, gender, sexual orientation, gender identity, age, physical or mental disability, genetic factors, military/veteran status or other characteristics protected by law.
09/27/2026
Full time
Job Description At Boeing, we innovate and collaborate to make the world a better place. We're committed to fostering an environment for every teammate that's welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us. Boeing Defense, Space & Security (BDS) is seeking a Lead Systems Engineer (Level 4) to support the Space Battle Management, Command and Control (SBMC2) programs in Colorado Springs, CO or Berkeley, MO. The Systems Engineer will perform as part of a high-performing team and have the opportunity to contribute to the development and delivery of innovative software solutions in an agile software development environment interfacing with internal and external stakeholders; to include a Joint Industry partnership Team (JIPT) and Working Groups, contributing to the design and development of next generation ground command and control capabilities. As a member of the Boeing team, you will be responsible for key portions of our development lifecycle, from idea creation and development, all the way through to maintenance and support of the customer's delivered system. More importantly, you will have the opportunity to make an impact on the results of our projects. We offer a collaborative mentoring environment where you have the opportunity to learn from others and be a mentor to others. Position Responsibilities Support definition of requirements, interfaces, and concept of operations through supporting and/or chairing working group meetings. Understand and communicate prioritization of efforts to multi-disciplinary team. Think abstractly and see the big picture while evaluating technical details Perform technical analyses to develop and validate models of system behavior; identify solutions to complex problems. Serve as a direct interface to both internal and external customers Conduct and support trade studies. Communicate across multiple functional groups (ex. Software, DevSecOps teams, hardware, cybersecurity, IT/infrastructure, networking, mission engineering, etc) Work in an industry consortium with diverse viewpoints Travel may be required up to 10% of the time; Domestically and/or internationally depending on business needs. This position requires an active U.S. Top Secret Security Clearance (U.S. Citizenship Required. (A U.S. Security Clearance that has been active in the past 24 months is considered active) Basic Qualifications (Required Skills/Experience) Bachelor of Science degree from an accredited course of study in engineering, engineering technology (includes manufacturing engineering technology), chemistry, physics, mathematics, data science, or computer science 9+ years of work-related engineering experience 3+ years of experience working on a software development effort, supporting the development of products from inception through delivery and operation (full product lifecycle) Experience with an interface standard, preferably Open Mission Systems (OMS), Universal Command and Control Interface (UCI) experience or Open Architecture standards Experience with Agile development Preferred Qualifications (Desired Skills/Experience) Satellite operations experience Familiarity with Space Domain Awareness Basic modeling skills with SysML, Enterprise Architect, or UML Knowledge of the space mission domain including launch vehicles, space systems, mission and space vehicle command & control (C2), specialized mission payloads, system integration, and interoperability. Conflict of Interest Successful candidates for this job must satisfy the Company's Conflict of Interest (COI) assessment process. Drug Free Workplace Boeing is a Drug Free Workplace where post offer applicants and employees are subject to testing for marijuana, cocaine, opioids, amphetamines, PCP, and alcohol when criteria is met as outlined in our policies. Employee Referral Referral to this job is eligible for bonus to qualifying candidates. Total Rewards At Boeing, we strive to deliver a Total Rewards package that will attract, engage and retain the top talent. Elements of the Total Rewards package include competitive base pay and variable compensation opportunities. The Boeing Company also provides eligible employees with an opportunity to enroll in a variety of benefit programs, generally including health insurance, flexible spending accounts, health savings accounts, retirement savings plans, life and disability insurance programs, and a number of programs that provide for both paid and unpaid time away from work. The specific programs and options available to any given employee may vary depending on eligibility factors such as geographic location, date of hire, and the applicability of collective bargaining agreements. The Boeing 401(k) helps you save for your future, with contributions from Boeing that can help you grow your retirement savings. Our best-in-class retirement benefit features: Best in class 401(k) plan: we'll match your contributions dollar for dollar, up to 10% of eligible pay with Immediate 100% vesting Student Loan Match: The Boeing 401(k) Student Loan Match allows eligible enrolled U.S. employees to have their qualified student loan debt payments counted, along with any match-eligible contributions they make, for purposes of determining the Company Match to employees' Boeing 401(k) accounts. Pay is based upon candidate experience and qualifications, as well as market and business considerations. Summary pay range: $136,850 - $185,150 Applications for this position will be accepted until Oct. 07, 2026 Export Control Requirements: This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a "U.S. Person" as defined by 22 C.F.R. 120.62 is required. "U.S. Person" includes U.S. Citizen, U.S. National, lawful permanent resident, refugee, or asylee. Export Control Details: US based job, US Person required Education Bachelor's Degree or Equivalent Required Relocation This position offers relocation based on candidate eligibility. Security Clearance This position requires an active U.S. Top Secret Security Clearance (U.S. Citizenship Required). (A U.S. Security Clearance that has been active in the past 24 months is considered active) Visa Sponsorship Employer will not sponsor applicants for employment visa status. Shift This position is for 1st shift Equal Opportunity Employer: Boeing is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, national origin, gender, sexual orientation, gender identity, age, physical or mental disability, genetic factors, military/veteran status or other characteristics protected by law.

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board