it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

38 jobs found

Email me jobs like this
Refine Search
Current Search
senior firmware engineer
Systems Architect - Audio Technology
Bose Corporation Remote, Oregon
At Bose Corporation, we believe sound is the most powerful force on earth - and for over 60 years, we have been a company built on innovation, excellence, and independence. Privately owned, fiercely customer-focused, and driven by our values, we continue to lead industries and transform lives through sound. Today, Bose Corporation is entering an exciting new era. Across multiple global Business Units and Global Functions, we are shaping the future of audio technology, automotive, luxury, and premium experiences. We invite you to join us in this transformation. Job Description Systems Architect-End-to-End Principal Systems Engineer Audio Technology About the Role The Audio Technology team is seeking a Principal Systems Engineer - Systems Architect, E2E to define complete end-to-end systems spanning hardware, software, audio signal processing, acoustics, connectivity, sensing, compute, and user experience. This is a senior technical leadership role for someone who thinks system-first. You will shape the overall architecture, go deep enough into individual domains to understand constraints and tradeoffs, and guide decisions across multiple engineering disciplines. Success requires broad technical depth, strong systems judgment, and the ability to earn credibility with specialists across hardware and software teams. What You'll Do Lead end-to-end system architecture for complex audio systems. Translate customer, product, and technology needs into coherent system architectures. Define system decomposition, interfaces, data flows, functional boundaries, and resource allocation. Drive tradeoffs across hardware, software, DSP, acoustics, connectivity, sensing, power, compute, performance, and cost. Identify cross-domain risks and dependencies early and drive system-level solutions. Use modeling, simulation, prototyping, and test data to evaluate architecture options. Develop reusable architectures and reference solutions that scale across applications. Provide Principal-level technical leadership across Audio Technology, suppliers, and technology partners. Minimum Qualifications Bachelor's degree in Systems Engineering, Electrical Engineering, Computer Engineering, Computer Science, Mechanical Engineering, or a related discipline, or equivalent practical experience. Significant experience developing and architecting complex multidisciplinary systems. Demonstrated experience leading end-to-end architecture across hardware and software. Strong understanding of embedded systems, including interactions among hardware, software, firmware, signal processing, and physical constraints. Experience defining system architectures, interfaces, functional decomposition, requirements, and technical tradeoffs. Demonstrated ability to work across engineering disciplines and optimize the overall system. Strong technical leadership and communication skills, including the ability to influence without direct authority. Preferred Qualifications Advanced degree in Systems Engineering or a related technical discipline. Experience architecting consumer audio, automotive audio, or other complex embedded systems. Strong knowledge of digital audio signal processing, acoustics, and audio system architectures. Experience with embedded processors, SoCs, DSPs, real-time operating systems, and Linux-based systems. Experience with distributed systems, connectivity, sensing, or hardware/software co-design. Experience developing reusable platforms, reference architectures, SDKs, frameworks, or technology building blocks. Experience working with semiconductor vendors, suppliers, automotive partners, or third-party integrators. Who You Are You are a systems thinker first. You start with the complete problem and intended experience, then determine how the hardware, software, and functional domains must work together to solve it. You can operate at the highest level of architecture while going deep enough into individual domains to understand what is possible, where the risks are, and which tradeoffs matter most. Most importantly, you want to help architect the future of Bose-turning advanced technology into coherent, scalable, differentiated systems that power exceptional audio experiences. At Bose, you're inspired to be and do your best and are rewarded for your unique talents! Our compensation is thoughtfully tailored to your skills, experience, education, and location, and goes beyond base salary. The hiring range for this position in the primary work location of Remote, California is: $168,300-$231,450.The hiring range for other Bose work locations may vary. In addition to competitive base pay we offer rewards including bonus programs, comprehensive health and welfare benefits, a 401(k) plan, plus exclusive perks designed to support your wellbeing, and a generous employee discount where you can immerse yourself in our products and experiences. We are a proudly independent company-driven by purpose, guided by our values, and united by a belief in the power of sound. As the world leader in audio experiences, we're creating what's next-pushing boundaries and delivering transformative sound experiences for people everywhere. Join us and make your next career move a mic-drop. Let's Make Waves.
09/24/2026
Full time
At Bose Corporation, we believe sound is the most powerful force on earth - and for over 60 years, we have been a company built on innovation, excellence, and independence. Privately owned, fiercely customer-focused, and driven by our values, we continue to lead industries and transform lives through sound. Today, Bose Corporation is entering an exciting new era. Across multiple global Business Units and Global Functions, we are shaping the future of audio technology, automotive, luxury, and premium experiences. We invite you to join us in this transformation. Job Description Systems Architect-End-to-End Principal Systems Engineer Audio Technology About the Role The Audio Technology team is seeking a Principal Systems Engineer - Systems Architect, E2E to define complete end-to-end systems spanning hardware, software, audio signal processing, acoustics, connectivity, sensing, compute, and user experience. This is a senior technical leadership role for someone who thinks system-first. You will shape the overall architecture, go deep enough into individual domains to understand constraints and tradeoffs, and guide decisions across multiple engineering disciplines. Success requires broad technical depth, strong systems judgment, and the ability to earn credibility with specialists across hardware and software teams. What You'll Do Lead end-to-end system architecture for complex audio systems. Translate customer, product, and technology needs into coherent system architectures. Define system decomposition, interfaces, data flows, functional boundaries, and resource allocation. Drive tradeoffs across hardware, software, DSP, acoustics, connectivity, sensing, power, compute, performance, and cost. Identify cross-domain risks and dependencies early and drive system-level solutions. Use modeling, simulation, prototyping, and test data to evaluate architecture options. Develop reusable architectures and reference solutions that scale across applications. Provide Principal-level technical leadership across Audio Technology, suppliers, and technology partners. Minimum Qualifications Bachelor's degree in Systems Engineering, Electrical Engineering, Computer Engineering, Computer Science, Mechanical Engineering, or a related discipline, or equivalent practical experience. Significant experience developing and architecting complex multidisciplinary systems. Demonstrated experience leading end-to-end architecture across hardware and software. Strong understanding of embedded systems, including interactions among hardware, software, firmware, signal processing, and physical constraints. Experience defining system architectures, interfaces, functional decomposition, requirements, and technical tradeoffs. Demonstrated ability to work across engineering disciplines and optimize the overall system. Strong technical leadership and communication skills, including the ability to influence without direct authority. Preferred Qualifications Advanced degree in Systems Engineering or a related technical discipline. Experience architecting consumer audio, automotive audio, or other complex embedded systems. Strong knowledge of digital audio signal processing, acoustics, and audio system architectures. Experience with embedded processors, SoCs, DSPs, real-time operating systems, and Linux-based systems. Experience with distributed systems, connectivity, sensing, or hardware/software co-design. Experience developing reusable platforms, reference architectures, SDKs, frameworks, or technology building blocks. Experience working with semiconductor vendors, suppliers, automotive partners, or third-party integrators. Who You Are You are a systems thinker first. You start with the complete problem and intended experience, then determine how the hardware, software, and functional domains must work together to solve it. You can operate at the highest level of architecture while going deep enough into individual domains to understand what is possible, where the risks are, and which tradeoffs matter most. Most importantly, you want to help architect the future of Bose-turning advanced technology into coherent, scalable, differentiated systems that power exceptional audio experiences. At Bose, you're inspired to be and do your best and are rewarded for your unique talents! Our compensation is thoughtfully tailored to your skills, experience, education, and location, and goes beyond base salary. The hiring range for this position in the primary work location of Remote, California is: $168,300-$231,450.The hiring range for other Bose work locations may vary. In addition to competitive base pay we offer rewards including bonus programs, comprehensive health and welfare benefits, a 401(k) plan, plus exclusive perks designed to support your wellbeing, and a generous employee discount where you can immerse yourself in our products and experiences. We are a proudly independent company-driven by purpose, guided by our values, and united by a belief in the power of sound. As the world leader in audio experiences, we're creating what's next-pushing boundaries and delivering transformative sound experiences for people everywhere. Join us and make your next career move a mic-drop. Let's Make Waves.
Systems Architect - Audio Technology
Bose Corporation
At Bose Corporation, we believe sound is the most powerful force on earth - and for over 60 years, we have been a company built on innovation, excellence, and independence. Privately owned, fiercely customer-focused, and driven by our values, we continue to lead industries and transform lives through sound. Today, Bose Corporation is entering an exciting new era. Across multiple global Business Units and Global Functions, we are shaping the future of audio technology, automotive, luxury, and premium experiences. We invite you to join us in this transformation. Job Description Systems Architect-End-to-End Principal Systems Engineer Audio Technology About the Role The Audio Technology team is seeking a Principal Systems Engineer - Systems Architect, E2E to define complete end-to-end systems spanning hardware, software, audio signal processing, acoustics, connectivity, sensing, compute, and user experience. This is a senior technical leadership role for someone who thinks system-first. You will shape the overall architecture, go deep enough into individual domains to understand constraints and tradeoffs, and guide decisions across multiple engineering disciplines. Success requires broad technical depth, strong systems judgment, and the ability to earn credibility with specialists across hardware and software teams. What You'll Do Lead end-to-end system architecture for complex audio systems. Translate customer, product, and technology needs into coherent system architectures. Define system decomposition, interfaces, data flows, functional boundaries, and resource allocation. Drive tradeoffs across hardware, software, DSP, acoustics, connectivity, sensing, power, compute, performance, and cost. Identify cross-domain risks and dependencies early and drive system-level solutions. Use modeling, simulation, prototyping, and test data to evaluate architecture options. Develop reusable architectures and reference solutions that scale across applications. Provide Principal-level technical leadership across Audio Technology, suppliers, and technology partners. Minimum Qualifications Bachelor's degree in Systems Engineering, Electrical Engineering, Computer Engineering, Computer Science, Mechanical Engineering, or a related discipline, or equivalent practical experience. Significant experience developing and architecting complex multidisciplinary systems. Demonstrated experience leading end-to-end architecture across hardware and software. Strong understanding of embedded systems, including interactions among hardware, software, firmware, signal processing, and physical constraints. Experience defining system architectures, interfaces, functional decomposition, requirements, and technical tradeoffs. Demonstrated ability to work across engineering disciplines and optimize the overall system. Strong technical leadership and communication skills, including the ability to influence without direct authority. Preferred Qualifications Advanced degree in Systems Engineering or a related technical discipline. Experience architecting consumer audio, automotive audio, or other complex embedded systems. Strong knowledge of digital audio signal processing, acoustics, and audio system architectures. Experience with embedded processors, SoCs, DSPs, real-time operating systems, and Linux-based systems. Experience with distributed systems, connectivity, sensing, or hardware/software co-design. Experience developing reusable platforms, reference architectures, SDKs, frameworks, or technology building blocks. Experience working with semiconductor vendors, suppliers, automotive partners, or third-party integrators. Who You Are You are a systems thinker first. You start with the complete problem and intended experience, then determine how the hardware, software, and functional domains must work together to solve it. You can operate at the highest level of architecture while going deep enough into individual domains to understand what is possible, where the risks are, and which tradeoffs matter most. Most importantly, you want to help architect the future of Bose-turning advanced technology into coherent, scalable, differentiated systems that power exceptional audio experiences. At Bose, you're inspired to be and do your best and are rewarded for your unique talents! Our compensation is thoughtfully tailored to your skills, experience, education, and location, and goes beyond base salary. The hiring range for this position in the primary work location of Remote, California is: $168,300-$231,450.The hiring range for other Bose work locations may vary. In addition to competitive base pay we offer rewards including bonus programs, comprehensive health and welfare benefits, a 401(k) plan, plus exclusive perks designed to support your wellbeing, and a generous employee discount where you can immerse yourself in our products and experiences. We are a proudly independent company-driven by purpose, guided by our values, and united by a belief in the power of sound. As the world leader in audio experiences, we're creating what's next-pushing boundaries and delivering transformative sound experiences for people everywhere. Join us and make your next career move a mic-drop. Let's Make Waves.
09/24/2026
Full time
At Bose Corporation, we believe sound is the most powerful force on earth - and for over 60 years, we have been a company built on innovation, excellence, and independence. Privately owned, fiercely customer-focused, and driven by our values, we continue to lead industries and transform lives through sound. Today, Bose Corporation is entering an exciting new era. Across multiple global Business Units and Global Functions, we are shaping the future of audio technology, automotive, luxury, and premium experiences. We invite you to join us in this transformation. Job Description Systems Architect-End-to-End Principal Systems Engineer Audio Technology About the Role The Audio Technology team is seeking a Principal Systems Engineer - Systems Architect, E2E to define complete end-to-end systems spanning hardware, software, audio signal processing, acoustics, connectivity, sensing, compute, and user experience. This is a senior technical leadership role for someone who thinks system-first. You will shape the overall architecture, go deep enough into individual domains to understand constraints and tradeoffs, and guide decisions across multiple engineering disciplines. Success requires broad technical depth, strong systems judgment, and the ability to earn credibility with specialists across hardware and software teams. What You'll Do Lead end-to-end system architecture for complex audio systems. Translate customer, product, and technology needs into coherent system architectures. Define system decomposition, interfaces, data flows, functional boundaries, and resource allocation. Drive tradeoffs across hardware, software, DSP, acoustics, connectivity, sensing, power, compute, performance, and cost. Identify cross-domain risks and dependencies early and drive system-level solutions. Use modeling, simulation, prototyping, and test data to evaluate architecture options. Develop reusable architectures and reference solutions that scale across applications. Provide Principal-level technical leadership across Audio Technology, suppliers, and technology partners. Minimum Qualifications Bachelor's degree in Systems Engineering, Electrical Engineering, Computer Engineering, Computer Science, Mechanical Engineering, or a related discipline, or equivalent practical experience. Significant experience developing and architecting complex multidisciplinary systems. Demonstrated experience leading end-to-end architecture across hardware and software. Strong understanding of embedded systems, including interactions among hardware, software, firmware, signal processing, and physical constraints. Experience defining system architectures, interfaces, functional decomposition, requirements, and technical tradeoffs. Demonstrated ability to work across engineering disciplines and optimize the overall system. Strong technical leadership and communication skills, including the ability to influence without direct authority. Preferred Qualifications Advanced degree in Systems Engineering or a related technical discipline. Experience architecting consumer audio, automotive audio, or other complex embedded systems. Strong knowledge of digital audio signal processing, acoustics, and audio system architectures. Experience with embedded processors, SoCs, DSPs, real-time operating systems, and Linux-based systems. Experience with distributed systems, connectivity, sensing, or hardware/software co-design. Experience developing reusable platforms, reference architectures, SDKs, frameworks, or technology building blocks. Experience working with semiconductor vendors, suppliers, automotive partners, or third-party integrators. Who You Are You are a systems thinker first. You start with the complete problem and intended experience, then determine how the hardware, software, and functional domains must work together to solve it. You can operate at the highest level of architecture while going deep enough into individual domains to understand what is possible, where the risks are, and which tradeoffs matter most. Most importantly, you want to help architect the future of Bose-turning advanced technology into coherent, scalable, differentiated systems that power exceptional audio experiences. At Bose, you're inspired to be and do your best and are rewarded for your unique talents! Our compensation is thoughtfully tailored to your skills, experience, education, and location, and goes beyond base salary. The hiring range for this position in the primary work location of Remote, California is: $168,300-$231,450.The hiring range for other Bose work locations may vary. In addition to competitive base pay we offer rewards including bonus programs, comprehensive health and welfare benefits, a 401(k) plan, plus exclusive perks designed to support your wellbeing, and a generous employee discount where you can immerse yourself in our products and experiences. We are a proudly independent company-driven by purpose, guided by our values, and united by a belief in the power of sound. As the world leader in audio experiences, we're creating what's next-pushing boundaries and delivering transformative sound experiences for people everywhere. Join us and make your next career move a mic-drop. Let's Make Waves.
Staff Design Verification Engineer, AI HW
Tenstorrent Austin, Texas
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Our Tensix team is building custom AI compute cores, RISC-V CPUs, and chiplet-based architectures for datacenter, edge, and automotive AI. Design Verification Engineers on this team validate compute IP and subsystems and build scalable DV infrastructure to keep verification fast, automated, and production-grade. This role is hybrid, based out of Toronto, ON, Austin, TX or Belgrade, Serbia. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Experienced in modern verification methodologies with strong SystemVerilog skills and exposure to structured testbench development. Comfortable working from block-level to system-level verification and reasoning about microarchitecture behavior from specs and waveforms. Proficient in Linux environments with Python, or Bash scripting for automate builds, parse logs, manage CI pipelines. Skilled in coverage-driven verification and confident debugging across RTL, testbench, and workload scenarios. Motivated by AI hardware and eager to learn verification strategies for new architectures and domains. What We Need Contribute to verification of Tensix IP and subsystems from early planning through tape-out, owning coverage goals and driving them to closure. Build and maintain reusable verification environments, agents, scoreboards, and supporting infrastructure in SystemVerilog and Python. Develop constrained-random and directed tests and close functional, code, and assertion coverage using data-driven approaches. Debug complex issues across RTL and testbench while collaborating with design, firmware, compiler, and architecture teams. Help evolve DV infrastructure including regression flows, CI/CD integration, EDA compute environments, and reproducible setups. Nice to have: Exposure to functional safety (FuSa) verification concepts for automotive AI workloads. Nice to have: Familiarity with formal verification methodologies, SVA (SystemVerilog Assertions), or formal property verification for complex control logic and microarchitectural corner cases. What You Will Learn How Tenstorrent verifies AI-native compute architectures including Tensix cores, RISC-V CPUs, matrix units, vector units, and datapath. How simulation, emulation, and software-driven workloads combine to validate both functional correctness and performance. How numerical precision formats such as BF16, FP4, and INT8 and on-chip dataflow impact verification complexity and coverage strategy. How to build scalable DV infrastructure across hybrid compute environments including containerized regression systems. How data analytics and modern AI tooling can enhance verification workflows and accelerate convergence. Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made. Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer. This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
09/24/2026
Full time
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Our Tensix team is building custom AI compute cores, RISC-V CPUs, and chiplet-based architectures for datacenter, edge, and automotive AI. Design Verification Engineers on this team validate compute IP and subsystems and build scalable DV infrastructure to keep verification fast, automated, and production-grade. This role is hybrid, based out of Toronto, ON, Austin, TX or Belgrade, Serbia. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Experienced in modern verification methodologies with strong SystemVerilog skills and exposure to structured testbench development. Comfortable working from block-level to system-level verification and reasoning about microarchitecture behavior from specs and waveforms. Proficient in Linux environments with Python, or Bash scripting for automate builds, parse logs, manage CI pipelines. Skilled in coverage-driven verification and confident debugging across RTL, testbench, and workload scenarios. Motivated by AI hardware and eager to learn verification strategies for new architectures and domains. What We Need Contribute to verification of Tensix IP and subsystems from early planning through tape-out, owning coverage goals and driving them to closure. Build and maintain reusable verification environments, agents, scoreboards, and supporting infrastructure in SystemVerilog and Python. Develop constrained-random and directed tests and close functional, code, and assertion coverage using data-driven approaches. Debug complex issues across RTL and testbench while collaborating with design, firmware, compiler, and architecture teams. Help evolve DV infrastructure including regression flows, CI/CD integration, EDA compute environments, and reproducible setups. Nice to have: Exposure to functional safety (FuSa) verification concepts for automotive AI workloads. Nice to have: Familiarity with formal verification methodologies, SVA (SystemVerilog Assertions), or formal property verification for complex control logic and microarchitectural corner cases. What You Will Learn How Tenstorrent verifies AI-native compute architectures including Tensix cores, RISC-V CPUs, matrix units, vector units, and datapath. How simulation, emulation, and software-driven workloads combine to validate both functional correctness and performance. How numerical precision formats such as BF16, FP4, and INT8 and on-chip dataflow impact verification complexity and coverage strategy. How to build scalable DV infrastructure across hybrid compute environments including containerized regression systems. How data analytics and modern AI tooling can enhance verification workflows and accelerate convergence. Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made. Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer. This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
Network Security Engineer
SpaceXAI Palo Alto, California
Job Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company's mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: We are seeking a seasoned Senior Network Security Engineer to join our dynamic security team. The ideal candidate will possess deep expertise in network security technologies, focusing on switching and routing systems within cloud-native and AI-focused infrastructure. You will thrive in a fast-paced, challenging environment, demonstrating adaptability, excellent multitasking skills, and the drive to tackle complex security issues independently while ensuring robust protection for our innovative technologies. RESPONSIBILITIES: Serve as a subject matter expert in network security, particularly firewalls, VPNs, IDS/IPS, routing protocols (e.g., BGP, OSPF), and switching technologies. Manage and update firewall configurations across our enterprise network to align with operational and security needs. Deploy new firewalls, switches, routers, and network security devices in response to evolving threats and demands. Develop and propose innovative network security solutions to address operational challenges in routing and switching environments. Enhance security processes through thorough documentation and change management. Act as the primary resolver for complex network security issues, including escalation support. Ensure network security systems, switches, and routers are up-to-date with patches, firmware, and maintenance. Monitor and respond to security events in cloud environments (e.g., AWS, GCP, Azure, Datacenter), with emphasis on network traffic analysis. BASIC QUALIFICATIONS: Bachelor's degree in Computer Science, Cybersecurity, Information Systems, or a related field. 4+ years of experience in network security engineering, with hands-on focus on switching and routing. Certifications like CISA, CRISC, CGEIT, Security+, CASP+, or similar preferred. Strong understanding of network security principles, protocols (e.g., TCP/IP, VLANs, ACLs), and best practices for secure routing and switching. Proficiency in at least one major cloud platform (AWS, GCP, or Azure) and its network security services (e.g., VPCs, Security Groups). Experience with network analysis tools such as Wireshark, tcpdump; and vendors including Cisco, Juniper, Palo Alto Networks. Familiarity with scripting languages (e.g., Python, Bash) for automation of network security tasks. PREFERRED SKILLS AND EXPERIENCE: Relevant network-specific certifications (e.g., CCNP Security, CCIE Security, JNCIP-SEC, PCNSE). Experience in multi-cloud environments and Infrastructure as Code tools like Terraform for network provisioning. Knowledge of DevSecOps practices tailored to network security integration. Experience building custom tools or integrations for enhancing network security operations. Interest in leveraging AI for network threat detection and automation. Contributions to open-source projects in network security or related tools. ITAR REQUIREMENTS: To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. 1157, or (iv) Asylee under 8 U.S.C. 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. COMPENSATION AND BENEFITS: $100,000 - $258,000 USD Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks. SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.
09/24/2026
Full time
Job Description Job Description SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company's mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates. ABOUT THE ROLE: We are seeking a seasoned Senior Network Security Engineer to join our dynamic security team. The ideal candidate will possess deep expertise in network security technologies, focusing on switching and routing systems within cloud-native and AI-focused infrastructure. You will thrive in a fast-paced, challenging environment, demonstrating adaptability, excellent multitasking skills, and the drive to tackle complex security issues independently while ensuring robust protection for our innovative technologies. RESPONSIBILITIES: Serve as a subject matter expert in network security, particularly firewalls, VPNs, IDS/IPS, routing protocols (e.g., BGP, OSPF), and switching technologies. Manage and update firewall configurations across our enterprise network to align with operational and security needs. Deploy new firewalls, switches, routers, and network security devices in response to evolving threats and demands. Develop and propose innovative network security solutions to address operational challenges in routing and switching environments. Enhance security processes through thorough documentation and change management. Act as the primary resolver for complex network security issues, including escalation support. Ensure network security systems, switches, and routers are up-to-date with patches, firmware, and maintenance. Monitor and respond to security events in cloud environments (e.g., AWS, GCP, Azure, Datacenter), with emphasis on network traffic analysis. BASIC QUALIFICATIONS: Bachelor's degree in Computer Science, Cybersecurity, Information Systems, or a related field. 4+ years of experience in network security engineering, with hands-on focus on switching and routing. Certifications like CISA, CRISC, CGEIT, Security+, CASP+, or similar preferred. Strong understanding of network security principles, protocols (e.g., TCP/IP, VLANs, ACLs), and best practices for secure routing and switching. Proficiency in at least one major cloud platform (AWS, GCP, or Azure) and its network security services (e.g., VPCs, Security Groups). Experience with network analysis tools such as Wireshark, tcpdump; and vendors including Cisco, Juniper, Palo Alto Networks. Familiarity with scripting languages (e.g., Python, Bash) for automation of network security tasks. PREFERRED SKILLS AND EXPERIENCE: Relevant network-specific certifications (e.g., CCNP Security, CCIE Security, JNCIP-SEC, PCNSE). Experience in multi-cloud environments and Infrastructure as Code tools like Terraform for network provisioning. Knowledge of DevSecOps practices tailored to network security integration. Experience building custom tools or integrations for enhancing network security operations. Interest in leveraging AI for network threat detection and automation. Contributions to open-source projects in network security or related tools. ITAR REQUIREMENTS: To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. 1157, or (iv) Asylee under 8 U.S.C. 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. COMPENSATION AND BENEFITS: $100,000 - $258,000 USD Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks. SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.
Senior Hardware Engineer
OPUS IVS INC Dexter, Michigan
Description: Company Overview At Opus IVS, our mission is to drive advancement in the automotive industry by assisting customers with complex vehicle repairs. Guided by our core values of Customer Focus, Innovation, Collaboration & Teamwork, and a Results-Driven approach, we continually strive to develop advanced technology that empowers us to fulfill our mission. Opus IVS technology & products has been a leader in the industry since the late 90's. Opus IVS offers modern collision shops an integrated platform of leading diagnostics and calibration solutions, anchored by expert technicians and cutting edge, patented technology. Position Summary The Senior Hardware Engineer will help develop state of the art vehicle diagnostic equipment and communication devices. This individual will work directly with Opus IVS's various technology departments to develop world-class diagnostic tools that meet and exceed our customers' needs. Responsibilities: Develop new embedded automotive diagnostic pass-thru products (PC boards and firmware). Build diagnostic systems including PC hardware, vehicle communication hardware, and battery management hardware. Develop new products/devices, features, and improve quality for new and existing hardware, firmware, and software. Improve code quality through writing unit tests, automation and performing code reviews. Design and develop product and protocol validation tests. Build product tools using C, C++, & C#. Test new designs and find solutions to issues. Create test systems for the new products. Implement new vehicle protocols and create test suites to confirm their operation. Interface with customers to collect information on the performance of our products. Other duties as assigned. Skills & Abilities : Customer Focus: Ability to understand and respond to the needs of customers with professionalism and care. Innovation: Ability to generate and apply creative ideas that improve work processes or add value. Collaboration & Teamwork: Ability to build cooperative relationships and contribute to group success. Results Driven: Ability to maintain a strong focus on achieving goals and delivering impactful results. Technical Aptitude: Ability to understand and use specific tools, systems, or technologies relevant to the role. Detail Oriented: Ability to focus on details, ensuring accuracy and precision in work tasks. Analytical Thinking: Ability to examine data and issues logically to draw insightful conclusions. Requirements: Qualifications: Bachelor's Degree in Hardware Engineering or related field; or equivalent experience plus education/training 7+ years' experience in hardware design or development Experience developing products using the ARM family of processors Strong understanding of electronic design for EMC / EMI compliance Strong knowledge of hardware design and components including MPUs, UART, CAN, LIN, A/D, D/A converters, RAM/Flash memories, and other digital electronics Experience with designing processors, power supplies, communication, analog, and digital electronic circuits Strong knowledge and experience of C, C++, and C# programming languages Ability to debug electronic circuits Ability to solder surface mount components onto a PC board Has knowledge about OBD (On Board Diagnostics) Experience with batteries and battery maintenance circuits Familiar with the SAE J2534 specification Some experience with USB, Bluetooth, and Wi-Fi connections Knowledgeable of Automotive diagnostic protocols Familiarity with Jira and Confluence WHAT WE OFFER: Competitive Pay: We know your value and we're not afraid to pay for it. We offer a competitive total compensation plan including salary, bonuses, tuition reimbursement, and a match contribution to your 401k. Time Off: Besides our competitive paid time off package, employees receive paid holidays and floating holidays. Benefits: We offer a comprehensive benefits package, including all the necessities such as medical, dental, and vision. Opportunity: to be a part of a fast-growing company working to make the world safer! We are an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, disability status or any other characteristic protected by state, federal, or local law. PHYSICAL DEMANDS The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. While performing the duties of the job, the employee is regularly required to use hands to finger, handle, or feel objects, tools or controls; reach with hands and arms; talk or hear. The employee frequently is required to stand, walk and sit. The employee is occasionally required to stoop, kneel, crouch or crawl. Specific vision abilities required by this job include close vision, color vision, peripheral vision, depth perception and the ability to adjust focus. The above information has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job. Compensation details: 00 Yearly Salary PI1a60726f76cb-0987
09/24/2026
Full time
Description: Company Overview At Opus IVS, our mission is to drive advancement in the automotive industry by assisting customers with complex vehicle repairs. Guided by our core values of Customer Focus, Innovation, Collaboration & Teamwork, and a Results-Driven approach, we continually strive to develop advanced technology that empowers us to fulfill our mission. Opus IVS technology & products has been a leader in the industry since the late 90's. Opus IVS offers modern collision shops an integrated platform of leading diagnostics and calibration solutions, anchored by expert technicians and cutting edge, patented technology. Position Summary The Senior Hardware Engineer will help develop state of the art vehicle diagnostic equipment and communication devices. This individual will work directly with Opus IVS's various technology departments to develop world-class diagnostic tools that meet and exceed our customers' needs. Responsibilities: Develop new embedded automotive diagnostic pass-thru products (PC boards and firmware). Build diagnostic systems including PC hardware, vehicle communication hardware, and battery management hardware. Develop new products/devices, features, and improve quality for new and existing hardware, firmware, and software. Improve code quality through writing unit tests, automation and performing code reviews. Design and develop product and protocol validation tests. Build product tools using C, C++, & C#. Test new designs and find solutions to issues. Create test systems for the new products. Implement new vehicle protocols and create test suites to confirm their operation. Interface with customers to collect information on the performance of our products. Other duties as assigned. Skills & Abilities : Customer Focus: Ability to understand and respond to the needs of customers with professionalism and care. Innovation: Ability to generate and apply creative ideas that improve work processes or add value. Collaboration & Teamwork: Ability to build cooperative relationships and contribute to group success. Results Driven: Ability to maintain a strong focus on achieving goals and delivering impactful results. Technical Aptitude: Ability to understand and use specific tools, systems, or technologies relevant to the role. Detail Oriented: Ability to focus on details, ensuring accuracy and precision in work tasks. Analytical Thinking: Ability to examine data and issues logically to draw insightful conclusions. Requirements: Qualifications: Bachelor's Degree in Hardware Engineering or related field; or equivalent experience plus education/training 7+ years' experience in hardware design or development Experience developing products using the ARM family of processors Strong understanding of electronic design for EMC / EMI compliance Strong knowledge of hardware design and components including MPUs, UART, CAN, LIN, A/D, D/A converters, RAM/Flash memories, and other digital electronics Experience with designing processors, power supplies, communication, analog, and digital electronic circuits Strong knowledge and experience of C, C++, and C# programming languages Ability to debug electronic circuits Ability to solder surface mount components onto a PC board Has knowledge about OBD (On Board Diagnostics) Experience with batteries and battery maintenance circuits Familiar with the SAE J2534 specification Some experience with USB, Bluetooth, and Wi-Fi connections Knowledgeable of Automotive diagnostic protocols Familiarity with Jira and Confluence WHAT WE OFFER: Competitive Pay: We know your value and we're not afraid to pay for it. We offer a competitive total compensation plan including salary, bonuses, tuition reimbursement, and a match contribution to your 401k. Time Off: Besides our competitive paid time off package, employees receive paid holidays and floating holidays. Benefits: We offer a comprehensive benefits package, including all the necessities such as medical, dental, and vision. Opportunity: to be a part of a fast-growing company working to make the world safer! We are an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, disability status or any other characteristic protected by state, federal, or local law. PHYSICAL DEMANDS The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. While performing the duties of the job, the employee is regularly required to use hands to finger, handle, or feel objects, tools or controls; reach with hands and arms; talk or hear. The employee frequently is required to stand, walk and sit. The employee is occasionally required to stoop, kneel, crouch or crawl. Specific vision abilities required by this job include close vision, color vision, peripheral vision, depth perception and the ability to adjust focus. The above information has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job. Compensation details: 00 Yearly Salary PI1a60726f76cb-0987
Senior Wireless Modem Software Engineer - 5G & Mobile Systems
Qualcomm San Diego, California
Qualcomm is seeking a Senior Engineer to design and optimize advanced wireless and media solutions for 5G, IoT, and mobile platforms. In this role, you will architect, implement, and debug high performance software and/or silicon features, analyze system performance, and drive end to end optimization across modem, RF, and application layers. You will collaborate closely with cross functional global teams to translate product requirements into scalable designs, conduct simulations, code reviews, and lab validation, and ensure carrier grade reliability. This position suits engineers who thrive in an innovative, fast paced environment and want to shape next generation connectivity. Responsibilities Design and implement advanced wireless and media features for 5 G, Io T, and mobile platforms Architect scalable software and/or silicon solutions aligned with product requirements Analyze system performance and drive end to end optimization across modem, RF, and application layers Debug complex issues using lab tools, logs, and simulations to ensure carrier grade reliability Collaborate with cross functional global teams including hardware, firmware, and systems engineering Conduct code reviews, design reviews, and contribute to best engineering practices Develop and run simulations, test plans, and validation procedures for new features Document designs, interfaces, and performance results for internal and external stakeholders Required Skills Wireless communications (4 G/5 G, LTE, NR) C/C++ programming Embedded systems development System level performance analysis Signal processing fundamentals So C and modem architecture Python or scripting for automation Lab debugging tools and test equipment Version control (e.g., Git) Simulation and modeling tools (e.g., MATLAB)
09/23/2026
Full time
Qualcomm is seeking a Senior Engineer to design and optimize advanced wireless and media solutions for 5G, IoT, and mobile platforms. In this role, you will architect, implement, and debug high performance software and/or silicon features, analyze system performance, and drive end to end optimization across modem, RF, and application layers. You will collaborate closely with cross functional global teams to translate product requirements into scalable designs, conduct simulations, code reviews, and lab validation, and ensure carrier grade reliability. This position suits engineers who thrive in an innovative, fast paced environment and want to shape next generation connectivity. Responsibilities Design and implement advanced wireless and media features for 5 G, Io T, and mobile platforms Architect scalable software and/or silicon solutions aligned with product requirements Analyze system performance and drive end to end optimization across modem, RF, and application layers Debug complex issues using lab tools, logs, and simulations to ensure carrier grade reliability Collaborate with cross functional global teams including hardware, firmware, and systems engineering Conduct code reviews, design reviews, and contribute to best engineering practices Develop and run simulations, test plans, and validation procedures for new features Document designs, interfaces, and performance results for internal and external stakeholders Required Skills Wireless communications (4 G/5 G, LTE, NR) C/C++ programming Embedded systems development System level performance analysis Signal processing fundamentals So C and modem architecture Python or scripting for automation Lab debugging tools and test equipment Version control (e.g., Git) Simulation and modeling tools (e.g., MATLAB)
Senior Hardware Engineer - $130000
OPUS IVS INC Dexter, Michigan
Description: Company Overview At Opus IVS, our mission is to drive advancement in the automotive industry by assisting customers with complex vehicle repairs. Guided by our core values of Customer Focus, Innovation, Collaboration & Teamwork, and a Results-Driven approach, we continually strive to develop advanced technology that empowers us to fulfill our mission. Opus IVS technology & products has been a leader in the industry since the late 90's. Opus IVS offers modern collision shops an integrated platform of leading diagnostics and calibration solutions, anchored by expert technicians and cutting edge, patented technology. Position Summary The Senior Hardware Engineer will help develop state of the art vehicle diagnostic equipment and communication devices. This individual will work directly with Opus IVS's various technology departments to develop world-class diagnostic tools that meet and exceed our customers' needs. Responsibilities: Develop new embedded automotive diagnostic pass-thru products (PC boards and firmware). Build diagnostic systems including PC hardware, vehicle communication hardware, and battery management hardware. Develop new products/devices, features, and improve quality for new and existing hardware, firmware, and software. Improve code quality through writing unit tests, automation and performing code reviews. Design and develop product and protocol validation tests. Build product tools using C, C++, & C#. Test new designs and find solutions to issues. Create test systems for the new products. Implement new vehicle protocols and create test suites to confirm their operation. Interface with customers to collect information on the performance of our products. Other duties as assigned. Skills & Abilities : Customer Focus: Ability to understand and respond to the needs of customers with professionalism and care. Innovation: Ability to generate and apply creative ideas that improve work processes or add value. Collaboration & Teamwork: Ability to build cooperative relationships and contribute to group success. Results Driven: Ability to maintain a strong focus on achieving goals and delivering impactful results. Technical Aptitude: Ability to understand and use specific tools, systems, or technologies relevant to the role. Detail Oriented: Ability to focus on details, ensuring accuracy and precision in work tasks. Analytical Thinking: Ability to examine data and issues logically to draw insightful conclusions. Requirements: Qualifications: Bachelor's Degree in Hardware Engineering or related field; or equivalent experience plus education/training 7+ years' experience in hardware design or development Experience developing products using the ARM family of processors Strong understanding of electronic design for EMC / EMI compliance Strong knowledge of hardware design and components including MPUs, UART, CAN, LIN, A/D, D/A converters, RAM/Flash memories, and other digital electronics Experience with designing processors, power supplies, communication, analog, and digital electronic circuits Strong knowledge and experience of C, C++, and C# programming languages Ability to debug electronic circuits Ability to solder surface mount components onto a PC board Has knowledge about OBD (On Board Diagnostics) Experience with batteries and battery maintenance circuits Familiar with the SAE J2534 specification Some experience with USB, Bluetooth, and Wi-Fi connections Knowledgeable of Automotive diagnostic protocols Familiarity with Jira and Confluence WHAT WE OFFER: Competitive Pay: We know your value and we're not afraid to pay for it. We offer a competitive total compensation plan including salary, bonuses, tuition reimbursement, and a match contribution to your 401k. Time Off: Besides our competitive paid time off package, employees receive paid holidays and floating holidays. Benefits: We offer a comprehensive benefits package, including all the necessities such as medical, dental, and vision. Opportunity: to be a part of a fast-growing company working to make the world safer! We are an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, disability status or any other characteristic protected by state, federal, or local law. PHYSICAL DEMANDS The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. While performing the duties of the job, the employee is regularly required to use hands to finger, handle, or feel objects, tools or controls; reach with hands and arms; talk or hear. The employee frequently is required to stand, walk and sit. The employee is occasionally required to stoop, kneel, crouch or crawl. Specific vision abilities required by this job include close vision, color vision, peripheral vision, depth perception and the ability to adjust focus. The above information has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job. Compensation details: 00 Yearly Salary PI1a60726f76cb-0987
09/23/2026
Full time
Description: Company Overview At Opus IVS, our mission is to drive advancement in the automotive industry by assisting customers with complex vehicle repairs. Guided by our core values of Customer Focus, Innovation, Collaboration & Teamwork, and a Results-Driven approach, we continually strive to develop advanced technology that empowers us to fulfill our mission. Opus IVS technology & products has been a leader in the industry since the late 90's. Opus IVS offers modern collision shops an integrated platform of leading diagnostics and calibration solutions, anchored by expert technicians and cutting edge, patented technology. Position Summary The Senior Hardware Engineer will help develop state of the art vehicle diagnostic equipment and communication devices. This individual will work directly with Opus IVS's various technology departments to develop world-class diagnostic tools that meet and exceed our customers' needs. Responsibilities: Develop new embedded automotive diagnostic pass-thru products (PC boards and firmware). Build diagnostic systems including PC hardware, vehicle communication hardware, and battery management hardware. Develop new products/devices, features, and improve quality for new and existing hardware, firmware, and software. Improve code quality through writing unit tests, automation and performing code reviews. Design and develop product and protocol validation tests. Build product tools using C, C++, & C#. Test new designs and find solutions to issues. Create test systems for the new products. Implement new vehicle protocols and create test suites to confirm their operation. Interface with customers to collect information on the performance of our products. Other duties as assigned. Skills & Abilities : Customer Focus: Ability to understand and respond to the needs of customers with professionalism and care. Innovation: Ability to generate and apply creative ideas that improve work processes or add value. Collaboration & Teamwork: Ability to build cooperative relationships and contribute to group success. Results Driven: Ability to maintain a strong focus on achieving goals and delivering impactful results. Technical Aptitude: Ability to understand and use specific tools, systems, or technologies relevant to the role. Detail Oriented: Ability to focus on details, ensuring accuracy and precision in work tasks. Analytical Thinking: Ability to examine data and issues logically to draw insightful conclusions. Requirements: Qualifications: Bachelor's Degree in Hardware Engineering or related field; or equivalent experience plus education/training 7+ years' experience in hardware design or development Experience developing products using the ARM family of processors Strong understanding of electronic design for EMC / EMI compliance Strong knowledge of hardware design and components including MPUs, UART, CAN, LIN, A/D, D/A converters, RAM/Flash memories, and other digital electronics Experience with designing processors, power supplies, communication, analog, and digital electronic circuits Strong knowledge and experience of C, C++, and C# programming languages Ability to debug electronic circuits Ability to solder surface mount components onto a PC board Has knowledge about OBD (On Board Diagnostics) Experience with batteries and battery maintenance circuits Familiar with the SAE J2534 specification Some experience with USB, Bluetooth, and Wi-Fi connections Knowledgeable of Automotive diagnostic protocols Familiarity with Jira and Confluence WHAT WE OFFER: Competitive Pay: We know your value and we're not afraid to pay for it. We offer a competitive total compensation plan including salary, bonuses, tuition reimbursement, and a match contribution to your 401k. Time Off: Besides our competitive paid time off package, employees receive paid holidays and floating holidays. Benefits: We offer a comprehensive benefits package, including all the necessities such as medical, dental, and vision. Opportunity: to be a part of a fast-growing company working to make the world safer! We are an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran, disability status or any other characteristic protected by state, federal, or local law. PHYSICAL DEMANDS The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions. While performing the duties of the job, the employee is regularly required to use hands to finger, handle, or feel objects, tools or controls; reach with hands and arms; talk or hear. The employee frequently is required to stand, walk and sit. The employee is occasionally required to stoop, kneel, crouch or crawl. Specific vision abilities required by this job include close vision, color vision, peripheral vision, depth perception and the ability to adjust focus. The above information has been designed to indicate the general nature and level of work performed by employees within this classification. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications required of employees assigned to this job. Compensation details: 00 Yearly Salary PI1a60726f76cb-0987
Sr. Engineer, SoC Design Verification
Tenstorrent Santa Clara, California
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a SoC Design Verification Engineer to lead pre-silicon verification of the Beowulf SoC, with focus on the Compute Subsystem (CSS), DDR memory subsystem, and Fabric NoC. This role will drive coverage, coherency, memory traffic, connectivity, error handling, and bring-up features critical to silicon success. This role is hybrid, based out of Boston, MA; Toronto, ON; or Santa Clara, CA. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Deeply curious about SoC architecture, compute systems, DDR behavior, and Fabric NoC/interconnect verification. Expert in UVM, SystemVerilog, coverage-driven verification, assertions, and subsystem-level debug. Experienced in verifying compute subsystems, DDR controllers and PHY-facing logic, NoC/interconnect protocols, coherency, ordering, and data movement. Comfortable with reset, power management, error handling, performance, and high-concurrency system scenarios. Proactive, detail-oriented, and effective in cross-functional technical discussions. Familiar with Python, C/C++, Tcl, CocoTB, or similar verification automation tools. What We Need Develop and own scalable verification environments for Beowulf SoC CSS, DDR, and Fabric NoC across subsystem and full-SoC levels. Write, refine, and execute test scenarios for compute operation, memory initialization and traffic, coherency, routing, ordering, QoS, power management, error handling, and data movement. Analyze coverage gaps, debug failures, drive root-cause analysis, and collaborate closely with architecture, RTL, firmware, emulation, and system teams. Drive verification planning, regression quality, coverage closure, and signoff for major SoC features. Automate verification flows using scripting, reusable infrastructure, and AI productivity tools. Mentor engineers and contribute to verification methodology and design-for-verification improvements. What You Will Learn In-depth SoC verification across compute, DDR, coherency, and Fabric NoC using modern workflows, tooling, and scripting. Integration of pre-silicon SoC verification with emulation, silicon bring-up, and platform validation strategies. How AI-driven automation reshapes modern DV workflows. Exposure to high-performance, system-level verification in advanced AI SoC designs. Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made. Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer. This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
09/23/2026
Full time
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a SoC Design Verification Engineer to lead pre-silicon verification of the Beowulf SoC, with focus on the Compute Subsystem (CSS), DDR memory subsystem, and Fabric NoC. This role will drive coverage, coherency, memory traffic, connectivity, error handling, and bring-up features critical to silicon success. This role is hybrid, based out of Boston, MA; Toronto, ON; or Santa Clara, CA. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Deeply curious about SoC architecture, compute systems, DDR behavior, and Fabric NoC/interconnect verification. Expert in UVM, SystemVerilog, coverage-driven verification, assertions, and subsystem-level debug. Experienced in verifying compute subsystems, DDR controllers and PHY-facing logic, NoC/interconnect protocols, coherency, ordering, and data movement. Comfortable with reset, power management, error handling, performance, and high-concurrency system scenarios. Proactive, detail-oriented, and effective in cross-functional technical discussions. Familiar with Python, C/C++, Tcl, CocoTB, or similar verification automation tools. What We Need Develop and own scalable verification environments for Beowulf SoC CSS, DDR, and Fabric NoC across subsystem and full-SoC levels. Write, refine, and execute test scenarios for compute operation, memory initialization and traffic, coherency, routing, ordering, QoS, power management, error handling, and data movement. Analyze coverage gaps, debug failures, drive root-cause analysis, and collaborate closely with architecture, RTL, firmware, emulation, and system teams. Drive verification planning, regression quality, coverage closure, and signoff for major SoC features. Automate verification flows using scripting, reusable infrastructure, and AI productivity tools. Mentor engineers and contribute to verification methodology and design-for-verification improvements. What You Will Learn In-depth SoC verification across compute, DDR, coherency, and Fabric NoC using modern workflows, tooling, and scripting. Integration of pre-silicon SoC verification with emulation, silicon bring-up, and platform validation strategies. How AI-driven automation reshapes modern DV workflows. Exposure to high-performance, system-level verification in advanced AI SoC designs. Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made. Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer. This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
Sr. Engineer, SoC Design Verification
Tenstorrent Boston, Massachusetts
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a SoC Design Verification Engineer to lead pre-silicon verification of the Beowulf SoC, with focus on the Compute Subsystem (CSS), DDR memory subsystem, and Fabric NoC. This role will drive coverage, coherency, memory traffic, connectivity, error handling, and bring-up features critical to silicon success. This role is hybrid, based out of Boston, MA; Toronto, ON; or Santa Clara, CA. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Deeply curious about SoC architecture, compute systems, DDR behavior, and Fabric NoC/interconnect verification. Expert in UVM, SystemVerilog, coverage-driven verification, assertions, and subsystem-level debug. Experienced in verifying compute subsystems, DDR controllers and PHY-facing logic, NoC/interconnect protocols, coherency, ordering, and data movement. Comfortable with reset, power management, error handling, performance, and high-concurrency system scenarios. Proactive, detail-oriented, and effective in cross-functional technical discussions. Familiar with Python, C/C++, Tcl, CocoTB, or similar verification automation tools. What We Need Develop and own scalable verification environments for Beowulf SoC CSS, DDR, and Fabric NoC across subsystem and full-SoC levels. Write, refine, and execute test scenarios for compute operation, memory initialization and traffic, coherency, routing, ordering, QoS, power management, error handling, and data movement. Analyze coverage gaps, debug failures, drive root-cause analysis, and collaborate closely with architecture, RTL, firmware, emulation, and system teams. Drive verification planning, regression quality, coverage closure, and signoff for major SoC features. Automate verification flows using scripting, reusable infrastructure, and AI productivity tools. Mentor engineers and contribute to verification methodology and design-for-verification improvements. What You Will Learn In-depth SoC verification across compute, DDR, coherency, and Fabric NoC using modern workflows, tooling, and scripting. Integration of pre-silicon SoC verification with emulation, silicon bring-up, and platform validation strategies. How AI-driven automation reshapes modern DV workflows. Exposure to high-performance, system-level verification in advanced AI SoC designs. Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made. Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer. This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
09/23/2026
Full time
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities. Tenstorrent is seeking a SoC Design Verification Engineer to lead pre-silicon verification of the Beowulf SoC, with focus on the Compute Subsystem (CSS), DDR memory subsystem, and Fabric NoC. This role will drive coverage, coherency, memory traffic, connectivity, error handling, and bring-up features critical to silicon success. This role is hybrid, based out of Boston, MA; Toronto, ON; or Santa Clara, CA. We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting. Who You Are Deeply curious about SoC architecture, compute systems, DDR behavior, and Fabric NoC/interconnect verification. Expert in UVM, SystemVerilog, coverage-driven verification, assertions, and subsystem-level debug. Experienced in verifying compute subsystems, DDR controllers and PHY-facing logic, NoC/interconnect protocols, coherency, ordering, and data movement. Comfortable with reset, power management, error handling, performance, and high-concurrency system scenarios. Proactive, detail-oriented, and effective in cross-functional technical discussions. Familiar with Python, C/C++, Tcl, CocoTB, or similar verification automation tools. What We Need Develop and own scalable verification environments for Beowulf SoC CSS, DDR, and Fabric NoC across subsystem and full-SoC levels. Write, refine, and execute test scenarios for compute operation, memory initialization and traffic, coherency, routing, ordering, QoS, power management, error handling, and data movement. Analyze coverage gaps, debug failures, drive root-cause analysis, and collaborate closely with architecture, RTL, firmware, emulation, and system teams. Drive verification planning, regression quality, coverage closure, and signoff for major SoC features. Automate verification flows using scripting, reusable infrastructure, and AI productivity tools. Mentor engineers and contribute to verification methodology and design-for-verification improvements. What You Will Learn In-depth SoC verification across compute, DDR, coherency, and Fabric NoC using modern workflows, tooling, and scripting. Integration of pre-silicon SoC verification with emulation, silicon bring-up, and platform validation strategies. How AI-driven automation reshapes modern DV workflows. Exposure to high-performance, system-level verification in advanced AI SoC designs. Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made. Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer. This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
Sr. Embedded Software Engineer
Stic Los Angeles, California
About Us Stic is modernizing out-of-home, one of advertising's oldest formats. Billboards and transit ads have always sold impressions on faith. Stic replaces faith by turning everyday vehicles into GPS-tracked, mobile media. Brands get proof instead of best guesses. Drivers earn passive income on miles they're already driving. Stic Vision is how it works. An edge AI device small enough to live inside a decal, streaming informative analytics with no vehicle-level integration required. The global out-of-home advertising market is worth roughly $40 billion a year. Stic is building the measurement layer that legacy OOH never had. Stic is trusted by brands like TikTok, Dunkin Donuts, e.l.f., and PacSun. The traction is already showing up in results: a SXSW campaign for TikTok Radio put 44 vehicles on the road in Austin and generated 3.9 million tracked impressions. That's real, revenue-generating proof in a category still mostly sold on assumptions. Stic is lean by design. People joining now build the playbook rather than inheriting one. This means fast execution, without the layers of process that slow down bigger companies. About You As a Senior Embedded Software Engineer at Stic, you will own the firmware and low-level software stack for the Stic Vision v3.0. A sticker-form-factor edge AI device that deploys on any vehicle, any window, any surface, delivering machine vision, real-time analytics, and AI inference at the edge without any vehicle-level integration. This role sits at the intersection of hardware and intelligence: you will write the firmware that drives computer vision, wireless communication, and sensor fusion on a compact, low-power device operating in real-world automotive environments. Stic's mission is to bring edge AI sensing to the 1.4 billion legacy vehicles that autonomous platforms cannot reach. This role is foundational to that mission. The firmware you write is what turns a manufacturable sticker into a living sensor node. Why This Role Matters This role is the backbone of Stic's transition from prototype to a fully integrated, proprietary hardware and intelligence platform. The firmware you write directly powers data capture, edge intelligence, and real-world deployment at scale across the largest untapped sensor network on the planet: the existing vehicle fleet. What You'll Do Develop and maintain embedded firmware in C/C++ for microcontrollers and wireless modules powering the Stic Vision v3.0 Own OpenMV-based computer vision firmware for person detection, vehicle counting, facial sentiment analysis, and demographic measurement Implement BLE communication stacks - advertising, scanning, data payloads, and RSSI-based logic Integrate cellular (LTE/5G), GNSS, IMU, camera, and sensor interfaces into a unified firmware architecture Build low-power firmware architectures including sleep states, duty cycling, and battery optimization for field-deployed devices Define and implement device-to-cloud and device-to-phone communication protocols Support OTA update pipelines, field diagnostics, and reliability improvements Collaborate with hardware, computer vision, and cloud teams to ensure seamless full-stack integration Contribute to board bring-up, hardware/firmware debugging, and system validation across EVT/DVT/PVT cycles What You Have to Have 3-5+ years of professional embedded software engineering experience Strong proficiency in C, micropython and C++ for embedded systems Hands-on experience with BLE firmware development (advertising, scanning, payload design, RSSI logic) Experience with real-time operating systems (RTOS), interrupt handling, and low-level memory management Familiarity with communication protocols including SPI, I2C, UART, and wireless stacks Proficiency with debugging tools including JTAG, oscilloscopes, and logic analyzers Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or equivalent experience What You Should Ideally Have Experience with OpenMV or similar embedded computer vision / ML-enabled camera modules Familiarity with nRF, STM32, Snapdragon, NVIDIA Jetson, ESP32, or similar platforms Experience with LTE/5G module integration and GNSS systems Background in edge AI inference, sensor fusion, or on-device ML Exposure to EVT/DVT/PVT processes and scaling hardware from prototype to production Experience in IoT, automotive, or rugged field-deployed systems Base Salary: $130,000 - $190,000 Work Location: Los Angeles - in person Benefits Join a well-funded early-stage company disrupting traditional advertising with patent pending technology 26 days off per year (15 PTO / 11 holidays) prorated for any partial year of employment Explosive career growth Health & dental insurance 401(k) Equal Employment Opportunity Stic is an equal opportunity employer committed to building a diverse, equitable, and inclusive workplace. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity or expression, pregnancy, age, national origin, ancestry, citizenship, disability, genetic information, medical condition, marital status, military or veteran status, or any other characteristic protected by federal, state, or local law. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training. Use of Artificial Intelligence and Automated Tools in Recruiting Stic uses artificial intelligence and automated tools in parts of our recruiting process. These tools may help organize and review applications, transcribe or summarize interviews, and assist our team in assessing job-related qualifications. Artificial intelligence does not make hiring decisions. A member of our team reviews candidate materials and makes every screening, interview, and hiring decision. We do not use these tools to evaluate candidates on the basis of race, color, national origin, ancestry, sex, gender, gender identity or expression, sexual orientation, religion, age, disability, medical condition, genetic information, marital status, military or veteran status, or any other characteristic protected by federal or California law. Reasonable Accommodations & Alternative Process Stic is committed to providing reasonable accommodations for qualified individuals with disabilities and individuals with sincerely held religious beliefs in our job application and interview procedures. If you need assistance or an accommodation due to a disability, wish to request an alternative to any AI-assisted step in our recruiting process, or have questions about how these tools are used, please contact us at . Requesting an accommodation or alternative will not negatively impact your application or candidacy. Privacy Notice at Collection Personal information we collect includes identifiers, biometric information, internet activity, geolocation data, audio, electronic, video, or image information, professional or employment-related information, and inferences drawn from the above information. We do not sell your personal information. We may share your personal information with service providers who assist with payment processing, data analytics, campaign management, and technology infrastructure, and with brand partners to the extent necessary to demonstrate campaign performance. We do not share your personal information for cross-context behavioral advertising.
09/23/2026
Full time
About Us Stic is modernizing out-of-home, one of advertising's oldest formats. Billboards and transit ads have always sold impressions on faith. Stic replaces faith by turning everyday vehicles into GPS-tracked, mobile media. Brands get proof instead of best guesses. Drivers earn passive income on miles they're already driving. Stic Vision is how it works. An edge AI device small enough to live inside a decal, streaming informative analytics with no vehicle-level integration required. The global out-of-home advertising market is worth roughly $40 billion a year. Stic is building the measurement layer that legacy OOH never had. Stic is trusted by brands like TikTok, Dunkin Donuts, e.l.f., and PacSun. The traction is already showing up in results: a SXSW campaign for TikTok Radio put 44 vehicles on the road in Austin and generated 3.9 million tracked impressions. That's real, revenue-generating proof in a category still mostly sold on assumptions. Stic is lean by design. People joining now build the playbook rather than inheriting one. This means fast execution, without the layers of process that slow down bigger companies. About You As a Senior Embedded Software Engineer at Stic, you will own the firmware and low-level software stack for the Stic Vision v3.0. A sticker-form-factor edge AI device that deploys on any vehicle, any window, any surface, delivering machine vision, real-time analytics, and AI inference at the edge without any vehicle-level integration. This role sits at the intersection of hardware and intelligence: you will write the firmware that drives computer vision, wireless communication, and sensor fusion on a compact, low-power device operating in real-world automotive environments. Stic's mission is to bring edge AI sensing to the 1.4 billion legacy vehicles that autonomous platforms cannot reach. This role is foundational to that mission. The firmware you write is what turns a manufacturable sticker into a living sensor node. Why This Role Matters This role is the backbone of Stic's transition from prototype to a fully integrated, proprietary hardware and intelligence platform. The firmware you write directly powers data capture, edge intelligence, and real-world deployment at scale across the largest untapped sensor network on the planet: the existing vehicle fleet. What You'll Do Develop and maintain embedded firmware in C/C++ for microcontrollers and wireless modules powering the Stic Vision v3.0 Own OpenMV-based computer vision firmware for person detection, vehicle counting, facial sentiment analysis, and demographic measurement Implement BLE communication stacks - advertising, scanning, data payloads, and RSSI-based logic Integrate cellular (LTE/5G), GNSS, IMU, camera, and sensor interfaces into a unified firmware architecture Build low-power firmware architectures including sleep states, duty cycling, and battery optimization for field-deployed devices Define and implement device-to-cloud and device-to-phone communication protocols Support OTA update pipelines, field diagnostics, and reliability improvements Collaborate with hardware, computer vision, and cloud teams to ensure seamless full-stack integration Contribute to board bring-up, hardware/firmware debugging, and system validation across EVT/DVT/PVT cycles What You Have to Have 3-5+ years of professional embedded software engineering experience Strong proficiency in C, micropython and C++ for embedded systems Hands-on experience with BLE firmware development (advertising, scanning, payload design, RSSI logic) Experience with real-time operating systems (RTOS), interrupt handling, and low-level memory management Familiarity with communication protocols including SPI, I2C, UART, and wireless stacks Proficiency with debugging tools including JTAG, oscilloscopes, and logic analyzers Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or equivalent experience What You Should Ideally Have Experience with OpenMV or similar embedded computer vision / ML-enabled camera modules Familiarity with nRF, STM32, Snapdragon, NVIDIA Jetson, ESP32, or similar platforms Experience with LTE/5G module integration and GNSS systems Background in edge AI inference, sensor fusion, or on-device ML Exposure to EVT/DVT/PVT processes and scaling hardware from prototype to production Experience in IoT, automotive, or rugged field-deployed systems Base Salary: $130,000 - $190,000 Work Location: Los Angeles - in person Benefits Join a well-funded early-stage company disrupting traditional advertising with patent pending technology 26 days off per year (15 PTO / 11 holidays) prorated for any partial year of employment Explosive career growth Health & dental insurance 401(k) Equal Employment Opportunity Stic is an equal opportunity employer committed to building a diverse, equitable, and inclusive workplace. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity or expression, pregnancy, age, national origin, ancestry, citizenship, disability, genetic information, medical condition, marital status, military or veteran status, or any other characteristic protected by federal, state, or local law. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training. Use of Artificial Intelligence and Automated Tools in Recruiting Stic uses artificial intelligence and automated tools in parts of our recruiting process. These tools may help organize and review applications, transcribe or summarize interviews, and assist our team in assessing job-related qualifications. Artificial intelligence does not make hiring decisions. A member of our team reviews candidate materials and makes every screening, interview, and hiring decision. We do not use these tools to evaluate candidates on the basis of race, color, national origin, ancestry, sex, gender, gender identity or expression, sexual orientation, religion, age, disability, medical condition, genetic information, marital status, military or veteran status, or any other characteristic protected by federal or California law. Reasonable Accommodations & Alternative Process Stic is committed to providing reasonable accommodations for qualified individuals with disabilities and individuals with sincerely held religious beliefs in our job application and interview procedures. If you need assistance or an accommodation due to a disability, wish to request an alternative to any AI-assisted step in our recruiting process, or have questions about how these tools are used, please contact us at . Requesting an accommodation or alternative will not negatively impact your application or candidacy. Privacy Notice at Collection Personal information we collect includes identifiers, biometric information, internet activity, geolocation data, audio, electronic, video, or image information, professional or employment-related information, and inferences drawn from the above information. We do not sell your personal information. We may share your personal information with service providers who assist with payment processing, data analytics, campaign management, and technology infrastructure, and with brand partners to the extent necessary to demonstrate campaign performance. We do not share your personal information for cross-context behavioral advertising.
Senior Operations Engineer, MetalDev
CoreWeave New York, New York
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
09/23/2026
Full time
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
Senior Operations Engineer, MetalDev
CoreWeave Livingston, New Jersey
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
09/23/2026
Full time
CoreWeave is The Essential Cloud for AI . Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . About the Role CoreWeave's MetalDev team is seeking an experienced Senior Operations Engineer to join the MetalDev Operations team. Reporting to the Engineering Manager, this hands-on technical role focuses on service reliability, observability, and operational excellence. You will develop deep expertise in the Redfish-based services and automation tools used by frontline operations teams during data center bring-ups and in production environments. You will identify gaps in tooling and operational processes, partner with the engineering team to develop and validate fixes, and help ensure our automation operates reliably at scale. You will also submit change requests to hardware and firmware vendors, validate vendor-provided fixes, and continuously improve runbooks and on-call documentation to reduce recurring incidents and operational toil. In this role, you will work with cutting-edge infrastructure, including high-performance NVIDIA GPU servers, Cooling Distribution Units (CDUs), NVLink switches, and power shelves supporting GB200 and GB300 Vera Rubin NVL72 systems and custom in-house hardware. You will collaborate closely with the Hardware Engineering, Fleet Operations, and the Service Engineering teams. Approximately 80% of the role will focus on day-to-day operations, production support, and incident response. The remaining 20% will focus on improving operational processes, observability, documentation, and remediation capabilities and automation to prevent recurring incidents and increase service reliability. Key Responsibilities Triage and Troubleshooting Troubleshoot the team owned services, including initialization and reboot issues. Diagnose problems involving BMCs of, servers, DPUs, power shelves, Cooling Distribution Units (CDUs). Monitor fleet health, identify unhealthy devices, and coordinate remediation using established tools and procedures. Partner with Fleet Operations and other engineering teams to resolve complex or recurring production issues. Perform root-cause analysis and post-incident reviews, ensuring corrective actions are documented, tracked, and completed. Maintain clear incident communications, runbooks, escalation procedures, and operational records. Support the team owned services in production environments and participate in the team's on-call rotation. Observability, Reliability, and Vendor Partnerships Own monitoring, dashboards, alerts, and operational KPIs for the team owned services using Prometheus and Grafana. Define reliability objectives and drive measurable reductions in incidents, escalations, and recurring support requests. Investigate hardware, firmware, and issues in partnership with internal engineering teams and external vendors. Manage vendor support cases involving BMCs, servers, power systems, and cooling infrastructure. Collect and provide diagnostic data, track issues through resolution, and validate vendor fixes before production rollout. Documentation and Continuous Improvement Create and maintain operational documentation, troubleshooting guides, escalation procedures, and service-support materials. Capture and share incident findings and operational knowledge across the team and partner teams. Identify repetitive operational tasks and develop more efficient, consistent remediation processes. Evaluate team processes using operational data and incident trends, and recommend improvements. Minimum Qualifications 5+ of experience in cloud operations, site reliability engineering (SRE), infrastructure operations, or a related technical field. Working knowledge of Kubernetes and at least one public cloud platform, such as AWS or GCP. Experience deploying and supporting containerized applications in Kubernetes environments. Experience with incident management practices, including incident response, escalation, and post-incident review processes. Experience using Prometheus, Grafana, and PromQL for monitoring, alerting, and troubleshooting. Strong knowledge of Linux system administration and internals and scripting. Experience troubleshooting complex issues across software services, operating systems, networks, and physical infrastructure. Experience participating in an on-call rotation supporting production services. Strong analytical and problem-solving skills, with a methodical approach to troubleshooting. Excellent written and verbal communication skills, particularly during high-impact incidents. Strong documentation skills and attention to detail. Preferred Qualifications Experience with server hardware, BMCs, Redfish, IPMI, or hardware-management services. Experience troubleshooting server provisioning, reboot, provisioning, or lifecycle-management failures. Familiarity with high-performance computing, GPU infrastructure, DPUs, or large-scale AI clusters. Experience working in data center environments, including server racks, power-distribution equipment, and cooling systems. Experience collaborating directly with hardware or firmware vendors to qualify and validate fixes. Bachelor's degree in computer science, engineering, or a related discipline-or equivalent practical experience. Understanding of Python or Golang. Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams - even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk. You enjoy working close to the hardware and are curious about how GPUs, servers, and data centers fit together. You thrive in infrastructure environments where reliability, performance, and automation matter as much as features. You like collaborating across hardware, platform, and product teams to solve complex, ambiguous problems. Why CoreWeave? At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: Be Curious at Your Core Act Like an Owner Empower Employees Deliver Best-in-Class Client Experiences Achieve More Together We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! The base salary range for this role is $134,000 to $179,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility). What We Offer The range we've posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location. In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include: Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insurance Flexible Spending Account Health Savings Account Tuition Reimbursement Ability to Participate in Employee Stock Purchase Program (ESPP) Mental Wellness Benefits through Spring Health Family-Forming support provided by Carrot Paid Parental Leave Flexible, full-service childcare support with Kinside 401(k) with a generous employer match Flexible PTO Catered lunch each day in our office and data center locations A casual work environment A work culture focused on innovative disruption California Applicants California Consumer Privacy Act Equal Opportunity & Accommodations CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace . click apply for full job details
Senior Embedded Firmware Engineer
Trackonomy San Jose, California
About Trackonomy Trackonomy is pioneering a transformative network of interconnected objects. Our goal is to bring inanimate objects to life, enabling them to communicate, think, and interact in real time. We are building the operating system for the connected world-one object and sensor at a time. Our customers span logistics, industrial, utilities, healthcare, and government sectors, using our solutions for predictive maintenance, workflow optimization, asset protection, safety, security, and environmental monitoring. Backed by leading investors including Kleiner Perkins and 8VC, Trackonomy is one of Silicon Valley's fastest-growing IoT companies. Our executive leadership team has worked together for more than two decades, holding key leadership positions at companies including Flextronics, Heptagon, GT-Nexus, and Digital Motors Corporation. Together, they have engineered breakthrough technology that is delivering extraordinary results for customers around the world. Are you ready to be part of the next chapter in Silicon Valley's hyper-growth IoT success story? The Role You will be a core engineer on our early-stage team. We work on everything from machine learning and security to high-performance computing, IoT devices, and dynamic web applications. Don't be surprised if you have the opportunity to touch nearly every system while working here, often collaborating with teammates on critical initiatives. Candidates must be comfortable working with microcontrollers and low-level hardware control in a test-driven development environment. The ideal candidate will have experience working with wireless communications modules, IoT technologies, RF protocols, and strong troubleshooting and prototyping skills. Experience Bachelor's or Master's degree in Electronics Engineering At least 6+ years of embedded software development with emphasis on C/C++ Background in real-time embedded systems Extensive work in firmware development, testing, and system-level bring-up and debugging Skilled in bench modifications and rapid prototyping of hardware/firmware solutions Strong interpersonal, organizational, and communication skills Effective collaborator who shares knowledge, learns from others, and supports cross-functional teams Self-starter with strong motivation and ownership mentality Preferred Skills Background in IoT systems and wireless/wired communication protocols, including BLE, LoRa, and LoRaWAN Skilled in firmware development on ARM-based microcontrollers and processors; familiarity with Nordic devices is a plus Knowledge of cellular protocols such as LTE, LTE-M/NB-IoT, and industrial wireless electronics Hands-on work with cellular modems and GPS/GNSS systems Ability to develop device drivers for sensors and communication modems Strong understanding of subsystem interfaces such as UART, I2C, SPI, and other standard chip-level protocols Proficiency in low-power embedded development, including performing power profiling on target devices Knowledge of firmware development best practices including testing, documentation, debugging, and code review Understanding of PCB design using schematic capture and layout tools (Eagle/Altium preferred) Capability to design bootloaders and implement firmware-over-the-air updates Familiarity with scripting languages (Python preferred) Ability to read and interpret complex electrical schematics Strong foundation in microcontrollers and embedded peripheral driver development Knowledge of embedded networking protocols such as RSTP, PTP, LLDP, and UDP/TCP is a plus Why Trackonomy? Meaningful Impact Trackonomy is dedicated to building technology that solves real-world problems. From fire prevention and environmental monitoring to safety, security, and operational efficiency, our innovations help organizations protect assets, improve outcomes, and save lives. Growth & Development At Trackonomy, your growth is driven by your capabilities, passion, ownership, and results. As part of an early-stage, high-growth company, you will have opportunities to take on diverse responsibilities, learn directly from experienced leaders, and expand your role as your impact grows. We believe talent and execution matter more than hierarchy. Team members are encouraged to explore new challenges, collaborate across functions, and continuously develop new skills. Ownership & Collaboration Our culture isn't something you join-it's something you help build. Every employee plays a meaningful role in our success and contributes to shaping our products, processes, and future. We value collaboration, knowledge sharing, accountability, and a strong sense of ownership. Benefits & Rewards Trackonomy understands that personal wellness is critical to a happy, healthy, and productive work environment. We offer: Platinum-level health benefits Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) Commuter benefits Employee Assistance Program (EAP) 401(k) plan Pre-IPO equity opportunity Regular performance reviews and feedback Learning and development opportunities for both individual contributors and aspiring leaders Equal Opportunity Employer Trackonomy Systems is proud to provide equal employment opportunities to all individuals regardless of race, color, religion, national origin, ancestry, physical or mental disability, sex, gender, gender identity, gender expression, sexual orientation, age, medical condition, genetic information, marital or registered domestic partnership status, military or veteran status, or any other characteristic protected by applicable law. We strive to provide a stellar experience throughout the application process and ensure all applicants receive fair consideration based solely on merit and business needs. The salary range for this role is $140,000 to $200,000, plus bonuses and Pre-IPO equity. It is uncommon for anyone to be hired at or near the top of the range. Final compensation is based on several factors, including skills, experience, training, business needs, cultural fit, level, and location. Trackonomy Systems is dedicated to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need assistance or accommodation during the interview process, please contact . When you apply to a job on this site, you acknowledge and agree that the personal data contained in your application will be collected and processed by Trackonomy Systems, Inc. and/or one of its subsidiaries ("Trackonomy") in accordance with our Applicant Privacy Notice. If you have any questions about our privacy practices, please contact .
09/23/2026
Full time
About Trackonomy Trackonomy is pioneering a transformative network of interconnected objects. Our goal is to bring inanimate objects to life, enabling them to communicate, think, and interact in real time. We are building the operating system for the connected world-one object and sensor at a time. Our customers span logistics, industrial, utilities, healthcare, and government sectors, using our solutions for predictive maintenance, workflow optimization, asset protection, safety, security, and environmental monitoring. Backed by leading investors including Kleiner Perkins and 8VC, Trackonomy is one of Silicon Valley's fastest-growing IoT companies. Our executive leadership team has worked together for more than two decades, holding key leadership positions at companies including Flextronics, Heptagon, GT-Nexus, and Digital Motors Corporation. Together, they have engineered breakthrough technology that is delivering extraordinary results for customers around the world. Are you ready to be part of the next chapter in Silicon Valley's hyper-growth IoT success story? The Role You will be a core engineer on our early-stage team. We work on everything from machine learning and security to high-performance computing, IoT devices, and dynamic web applications. Don't be surprised if you have the opportunity to touch nearly every system while working here, often collaborating with teammates on critical initiatives. Candidates must be comfortable working with microcontrollers and low-level hardware control in a test-driven development environment. The ideal candidate will have experience working with wireless communications modules, IoT technologies, RF protocols, and strong troubleshooting and prototyping skills. Experience Bachelor's or Master's degree in Electronics Engineering At least 6+ years of embedded software development with emphasis on C/C++ Background in real-time embedded systems Extensive work in firmware development, testing, and system-level bring-up and debugging Skilled in bench modifications and rapid prototyping of hardware/firmware solutions Strong interpersonal, organizational, and communication skills Effective collaborator who shares knowledge, learns from others, and supports cross-functional teams Self-starter with strong motivation and ownership mentality Preferred Skills Background in IoT systems and wireless/wired communication protocols, including BLE, LoRa, and LoRaWAN Skilled in firmware development on ARM-based microcontrollers and processors; familiarity with Nordic devices is a plus Knowledge of cellular protocols such as LTE, LTE-M/NB-IoT, and industrial wireless electronics Hands-on work with cellular modems and GPS/GNSS systems Ability to develop device drivers for sensors and communication modems Strong understanding of subsystem interfaces such as UART, I2C, SPI, and other standard chip-level protocols Proficiency in low-power embedded development, including performing power profiling on target devices Knowledge of firmware development best practices including testing, documentation, debugging, and code review Understanding of PCB design using schematic capture and layout tools (Eagle/Altium preferred) Capability to design bootloaders and implement firmware-over-the-air updates Familiarity with scripting languages (Python preferred) Ability to read and interpret complex electrical schematics Strong foundation in microcontrollers and embedded peripheral driver development Knowledge of embedded networking protocols such as RSTP, PTP, LLDP, and UDP/TCP is a plus Why Trackonomy? Meaningful Impact Trackonomy is dedicated to building technology that solves real-world problems. From fire prevention and environmental monitoring to safety, security, and operational efficiency, our innovations help organizations protect assets, improve outcomes, and save lives. Growth & Development At Trackonomy, your growth is driven by your capabilities, passion, ownership, and results. As part of an early-stage, high-growth company, you will have opportunities to take on diverse responsibilities, learn directly from experienced leaders, and expand your role as your impact grows. We believe talent and execution matter more than hierarchy. Team members are encouraged to explore new challenges, collaborate across functions, and continuously develop new skills. Ownership & Collaboration Our culture isn't something you join-it's something you help build. Every employee plays a meaningful role in our success and contributes to shaping our products, processes, and future. We value collaboration, knowledge sharing, accountability, and a strong sense of ownership. Benefits & Rewards Trackonomy understands that personal wellness is critical to a happy, healthy, and productive work environment. We offer: Platinum-level health benefits Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) Commuter benefits Employee Assistance Program (EAP) 401(k) plan Pre-IPO equity opportunity Regular performance reviews and feedback Learning and development opportunities for both individual contributors and aspiring leaders Equal Opportunity Employer Trackonomy Systems is proud to provide equal employment opportunities to all individuals regardless of race, color, religion, national origin, ancestry, physical or mental disability, sex, gender, gender identity, gender expression, sexual orientation, age, medical condition, genetic information, marital or registered domestic partnership status, military or veteran status, or any other characteristic protected by applicable law. We strive to provide a stellar experience throughout the application process and ensure all applicants receive fair consideration based solely on merit and business needs. The salary range for this role is $140,000 to $200,000, plus bonuses and Pre-IPO equity. It is uncommon for anyone to be hired at or near the top of the range. Final compensation is based on several factors, including skills, experience, training, business needs, cultural fit, level, and location. Trackonomy Systems is dedicated to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need assistance or accommodation during the interview process, please contact . When you apply to a job on this site, you acknowledge and agree that the personal data contained in your application will be collected and processed by Trackonomy Systems, Inc. and/or one of its subsidiaries ("Trackonomy") in accordance with our Applicant Privacy Notice. If you have any questions about our privacy practices, please contact .
Sr Product Security Engineer - Devices
Altice USA
Are you looking to Optimize your life? Start your exciting path to a rewarding career today! We are Optimum, a leader in the fast-paced world of connectivity, and we're seeking driven and enthusiastic professionals to join our team, empower lives, fuel businesses, and drive innovation. Connectivity is now longer a luxury, but a necessity. A career at Optimum means you'll be enabling progress and enhancing lives by providing reliable, high-speed connectivity solutions that keep the world connected. Our successes, now and in the future, are powered by our amazing product, a commitment to our people and culture, and the connections we make in our communities. If you are resourceful, collaborative, and passionate about delivering consistent excellence, Optimum is for you! Job Summary Optimum is a leading provider of Mobile, Broadband (DOCSIS, Fiber) and Video services in the United States for Business and Residential Customers. We are looking for a passionate and engaged individual who is willing to take on the challenge of scaling product security at a Fortune 500 company and building the cyber security foundation that will allow our development teams to Go Fast! The Product Security organization helps Optimum move faster, securely. We're a team of engineers who work to enable other teams to build products as quickly as possible while continuing to protect our customers. We support developers in shipping secure code by building security tools and services, providing security training and expertise, and advocating for best practices in authentication, authorization, and safe data handling across the company. As a Product Security Engineer focusing on embedded systems security for video, broadband, and Wi-Fi products, you'll be a trusted partner, collaborating closely with engineering and product teams to ensure security is a cornerstone of every product. You will partner with leadership to shape product strategy, advocate for strong security controls, and influence future product iterations, and serve as an industry expert in device security engineering practices and standards. By leveraging your deep industry knowledge, you'll lead the charge in implementing secure architecture and design principles, ensuring early detection and prevention of vulnerabilities. Your expertise in security assessments and software engineering will help identify and mitigate potential threats, while your mentorship and training efforts will foster a security-first culture. Responsibilities Collaborate with engineering and product teams to integrate security and secure-by-default guardrails into the product lifecycle, ensuring that security is a core consideration in all design and development decisions. Conduct Threat Modeling and Risk Assessments from the early stages of the product development lifecycle to identify, assess, and prioritize security risks, enabling proactive mitigation strategies. Perform rigorous security testing and reviews to uncover and address security weaknesses. Lead initiatives automating security processes from the developer workstation to cloud, SaaS, and datacenter environments. Foster a security-first culture by educating and empowering engineering and product teams through training, awareness campaigns, and mentorship, cultivating a strong security mindset. Stay updated on the latest security threats, vulnerabilities, and technology trends, and proactively implement improvements. Contribute to incident response efforts, investigate root causes, and implement corrective actions to minimize impact and prevent future occurrences. Conduct penetration testing, reverse engineering, and vulnerability research against firmware, hardware interfaces, and device software to surface real-world attack paths before adversaries do. Establish and operate Software Bill of Materials (SBOM) management, CVE triage, and coordinated vulnerability disclosure processes for connected devices throughout their lifecycle. Define and uphold secure-by-design requirements aligned to recognized IoT and embedded security standards (e.g., NIST IR 8259, ETSI EN , OWASP IoT Top 10). Qualifications Bachelor's degree in Computer Science, Electrical Engineering, or a related field. Master's degree is a plus. 5+ years of hands-on experience in software engineering and designing and delivering security-critical systems for internet-connected embedded devices. Proven expertise in embedded systems and product security, with a strong understanding of modern software development processes and methods, security best practices, threat modeling, and risk assessment. Excellent communication skills, both written and verbal, and the ability to communicate complex security concepts to technical and non-technical audiences, including senior leadership. Proven ability to establish credibility and build trust with engineers and operational staff. Expertise in conducting comprehensive threat modeling, risk assessments, and code reviews to identify and mitigate vulnerabilities. Experience utilizing and securing AI/ML models and AI-integrated solutions, a general understanding of AI concepts, AI governance and risk management, and a willingness to learn more. Proficiency in secure SDLC practices and practical experience with CI/CD pipelines and DevOps tools. Experience overseeing vulnerability and threat management at the platform and device levels. Strong understanding of cryptography and key management use cases. In-depth knowledge of networking protocols, peripheral and firmware security, secure boot, embedded Linux security, Android or iOS security, and PKI. Experience working with special purpose security hardware such as Trusted Platform Modules (TPMs) and Hardware Security Modules (HSMs). Proficiency in C and C++ for embedded software development and one or more modern programming languages like Golang, Python, Node, and Java. Hands-on experience with hardware-level security testing techniques, including JTAG/SWD, UART, SPI/I2C debug interfaces, firmware extraction, fault injection, and side-channel analysis. Working knowledge of reverse engineering and analysis tooling such as Ghidra, IDA Pro, Binary Ninja, binwalk, QEMU, and common debuggers across ARM, MIPS, and RISC-V architectures. At Optimum, every action and interaction we take part in, is driven by our three Guiding Principles: Do What's Right, Drive One Optimum, and Make It Happen. These aren't just words, they help us build trust, create real community, and embrace new ways of thinking. Our employees are empowered to do the right thing for our customers and co-workers and to recognize and reward these behaviors when we see them. It's all part of the bigger picture of "Be The Difference" where each employee knows they have the power to enact real change, share new ideas, and understand that learning never stop. If you have the drive to succeed and are ready to embark on a thrilling career, seize this opportunity today, and join our winning team. Together, we'll shape the future of connectivity. All job descriptions and required skills, qualifications and responsibilities for a particular position are subject to modification by the Company from time to time, in the Company's discretion based on business necessity. We are an Equal Opportunity Employer committed to recruiting, hiring and promoting qualified people of all backgrounds regardless of gender, race, color, creed, national origin, religion, age, marital status, pregnancy, physical or mental disability, sexual orientation, gender identity, military or veteran status, or any other basis protected by federal, state, or local law. The Company collects personal information about its applicants for employment that may include personal identifiers, professional or employment related information, photos, education information and/or protected classifications under federal and state law. This information is collected for employment purposes, including identification, work authorization, FCRA-compliant background screening, human resource administration and compliance with federal, state and local law. Applicants for employment with The Company will never be asked to provide money (even if reimbursable) as part of the job application or hiring process. Please review our Fraud FAQ for further details. We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.
09/23/2026
Full time
Are you looking to Optimize your life? Start your exciting path to a rewarding career today! We are Optimum, a leader in the fast-paced world of connectivity, and we're seeking driven and enthusiastic professionals to join our team, empower lives, fuel businesses, and drive innovation. Connectivity is now longer a luxury, but a necessity. A career at Optimum means you'll be enabling progress and enhancing lives by providing reliable, high-speed connectivity solutions that keep the world connected. Our successes, now and in the future, are powered by our amazing product, a commitment to our people and culture, and the connections we make in our communities. If you are resourceful, collaborative, and passionate about delivering consistent excellence, Optimum is for you! Job Summary Optimum is a leading provider of Mobile, Broadband (DOCSIS, Fiber) and Video services in the United States for Business and Residential Customers. We are looking for a passionate and engaged individual who is willing to take on the challenge of scaling product security at a Fortune 500 company and building the cyber security foundation that will allow our development teams to Go Fast! The Product Security organization helps Optimum move faster, securely. We're a team of engineers who work to enable other teams to build products as quickly as possible while continuing to protect our customers. We support developers in shipping secure code by building security tools and services, providing security training and expertise, and advocating for best practices in authentication, authorization, and safe data handling across the company. As a Product Security Engineer focusing on embedded systems security for video, broadband, and Wi-Fi products, you'll be a trusted partner, collaborating closely with engineering and product teams to ensure security is a cornerstone of every product. You will partner with leadership to shape product strategy, advocate for strong security controls, and influence future product iterations, and serve as an industry expert in device security engineering practices and standards. By leveraging your deep industry knowledge, you'll lead the charge in implementing secure architecture and design principles, ensuring early detection and prevention of vulnerabilities. Your expertise in security assessments and software engineering will help identify and mitigate potential threats, while your mentorship and training efforts will foster a security-first culture. Responsibilities Collaborate with engineering and product teams to integrate security and secure-by-default guardrails into the product lifecycle, ensuring that security is a core consideration in all design and development decisions. Conduct Threat Modeling and Risk Assessments from the early stages of the product development lifecycle to identify, assess, and prioritize security risks, enabling proactive mitigation strategies. Perform rigorous security testing and reviews to uncover and address security weaknesses. Lead initiatives automating security processes from the developer workstation to cloud, SaaS, and datacenter environments. Foster a security-first culture by educating and empowering engineering and product teams through training, awareness campaigns, and mentorship, cultivating a strong security mindset. Stay updated on the latest security threats, vulnerabilities, and technology trends, and proactively implement improvements. Contribute to incident response efforts, investigate root causes, and implement corrective actions to minimize impact and prevent future occurrences. Conduct penetration testing, reverse engineering, and vulnerability research against firmware, hardware interfaces, and device software to surface real-world attack paths before adversaries do. Establish and operate Software Bill of Materials (SBOM) management, CVE triage, and coordinated vulnerability disclosure processes for connected devices throughout their lifecycle. Define and uphold secure-by-design requirements aligned to recognized IoT and embedded security standards (e.g., NIST IR 8259, ETSI EN , OWASP IoT Top 10). Qualifications Bachelor's degree in Computer Science, Electrical Engineering, or a related field. Master's degree is a plus. 5+ years of hands-on experience in software engineering and designing and delivering security-critical systems for internet-connected embedded devices. Proven expertise in embedded systems and product security, with a strong understanding of modern software development processes and methods, security best practices, threat modeling, and risk assessment. Excellent communication skills, both written and verbal, and the ability to communicate complex security concepts to technical and non-technical audiences, including senior leadership. Proven ability to establish credibility and build trust with engineers and operational staff. Expertise in conducting comprehensive threat modeling, risk assessments, and code reviews to identify and mitigate vulnerabilities. Experience utilizing and securing AI/ML models and AI-integrated solutions, a general understanding of AI concepts, AI governance and risk management, and a willingness to learn more. Proficiency in secure SDLC practices and practical experience with CI/CD pipelines and DevOps tools. Experience overseeing vulnerability and threat management at the platform and device levels. Strong understanding of cryptography and key management use cases. In-depth knowledge of networking protocols, peripheral and firmware security, secure boot, embedded Linux security, Android or iOS security, and PKI. Experience working with special purpose security hardware such as Trusted Platform Modules (TPMs) and Hardware Security Modules (HSMs). Proficiency in C and C++ for embedded software development and one or more modern programming languages like Golang, Python, Node, and Java. Hands-on experience with hardware-level security testing techniques, including JTAG/SWD, UART, SPI/I2C debug interfaces, firmware extraction, fault injection, and side-channel analysis. Working knowledge of reverse engineering and analysis tooling such as Ghidra, IDA Pro, Binary Ninja, binwalk, QEMU, and common debuggers across ARM, MIPS, and RISC-V architectures. At Optimum, every action and interaction we take part in, is driven by our three Guiding Principles: Do What's Right, Drive One Optimum, and Make It Happen. These aren't just words, they help us build trust, create real community, and embrace new ways of thinking. Our employees are empowered to do the right thing for our customers and co-workers and to recognize and reward these behaviors when we see them. It's all part of the bigger picture of "Be The Difference" where each employee knows they have the power to enact real change, share new ideas, and understand that learning never stop. If you have the drive to succeed and are ready to embark on a thrilling career, seize this opportunity today, and join our winning team. Together, we'll shape the future of connectivity. All job descriptions and required skills, qualifications and responsibilities for a particular position are subject to modification by the Company from time to time, in the Company's discretion based on business necessity. We are an Equal Opportunity Employer committed to recruiting, hiring and promoting qualified people of all backgrounds regardless of gender, race, color, creed, national origin, religion, age, marital status, pregnancy, physical or mental disability, sexual orientation, gender identity, military or veteran status, or any other basis protected by federal, state, or local law. The Company collects personal information about its applicants for employment that may include personal identifiers, professional or employment related information, photos, education information and/or protected classifications under federal and state law. This information is collected for employment purposes, including identification, work authorization, FCRA-compliant background screening, human resource administration and compliance with federal, state and local law. Applicants for employment with The Company will never be asked to provide money (even if reimbursable) as part of the job application or hiring process. Please review our Fraud FAQ for further details. We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.
Sr Product Security Engineer - Devices
Altice USA Plano, Texas
Are you looking to Optimize your life? Start your exciting path to a rewarding career today! We are Optimum, a leader in the fast-paced world of connectivity, and we're seeking driven and enthusiastic professionals to join our team, empower lives, fuel businesses, and drive innovation. Connectivity is now longer a luxury, but a necessity. A career at Optimum means you'll be enabling progress and enhancing lives by providing reliable, high-speed connectivity solutions that keep the world connected. Our successes, now and in the future, are powered by our amazing product, a commitment to our people and culture, and the connections we make in our communities. If you are resourceful, collaborative, and passionate about delivering consistent excellence, Optimum is for you! Job Summary Optimum is a leading provider of Mobile, Broadband (DOCSIS, Fiber) and Video services in the United States for Business and Residential Customers. We are looking for a passionate and engaged individual who is willing to take on the challenge of scaling product security at a Fortune 500 company and building the cyber security foundation that will allow our development teams to Go Fast! The Product Security organization helps Optimum move faster, securely. We're a team of engineers who work to enable other teams to build products as quickly as possible while continuing to protect our customers. We support developers in shipping secure code by building security tools and services, providing security training and expertise, and advocating for best practices in authentication, authorization, and safe data handling across the company. As a Product Security Engineer focusing on embedded systems security for video, broadband, and Wi-Fi products, you'll be a trusted partner, collaborating closely with engineering and product teams to ensure security is a cornerstone of every product. You will partner with leadership to shape product strategy, advocate for strong security controls, and influence future product iterations, and serve as an industry expert in device security engineering practices and standards. By leveraging your deep industry knowledge, you'll lead the charge in implementing secure architecture and design principles, ensuring early detection and prevention of vulnerabilities. Your expertise in security assessments and software engineering will help identify and mitigate potential threats, while your mentorship and training efforts will foster a security-first culture. Responsibilities Collaborate with engineering and product teams to integrate security and secure-by-default guardrails into the product lifecycle, ensuring that security is a core consideration in all design and development decisions. Conduct Threat Modeling and Risk Assessments from the early stages of the product development lifecycle to identify, assess, and prioritize security risks, enabling proactive mitigation strategies. Perform rigorous security testing and reviews to uncover and address security weaknesses. Lead initiatives automating security processes from the developer workstation to cloud, SaaS, and datacenter environments. Foster a security-first culture by educating and empowering engineering and product teams through training, awareness campaigns, and mentorship, cultivating a strong security mindset. Stay updated on the latest security threats, vulnerabilities, and technology trends, and proactively implement improvements. Contribute to incident response efforts, investigate root causes, and implement corrective actions to minimize impact and prevent future occurrences. Conduct penetration testing, reverse engineering, and vulnerability research against firmware, hardware interfaces, and device software to surface real-world attack paths before adversaries do. Establish and operate Software Bill of Materials (SBOM) management, CVE triage, and coordinated vulnerability disclosure processes for connected devices throughout their lifecycle. Define and uphold secure-by-design requirements aligned to recognized IoT and embedded security standards (e.g., NIST IR 8259, ETSI EN , OWASP IoT Top 10). Qualifications Bachelor's degree in Computer Science, Electrical Engineering, or a related field. Master's degree is a plus. 5+ years of hands-on experience in software engineering and designing and delivering security-critical systems for internet-connected embedded devices. Proven expertise in embedded systems and product security, with a strong understanding of modern software development processes and methods, security best practices, threat modeling, and risk assessment. Excellent communication skills, both written and verbal, and the ability to communicate complex security concepts to technical and non-technical audiences, including senior leadership. Proven ability to establish credibility and build trust with engineers and operational staff. Expertise in conducting comprehensive threat modeling, risk assessments, and code reviews to identify and mitigate vulnerabilities. Experience utilizing and securing AI/ML models and AI-integrated solutions, a general understanding of AI concepts, AI governance and risk management, and a willingness to learn more. Proficiency in secure SDLC practices and practical experience with CI/CD pipelines and DevOps tools. Experience overseeing vulnerability and threat management at the platform and device levels. Strong understanding of cryptography and key management use cases. In-depth knowledge of networking protocols, peripheral and firmware security, secure boot, embedded Linux security, Android or iOS security, and PKI. Experience working with special purpose security hardware such as Trusted Platform Modules (TPMs) and Hardware Security Modules (HSMs). Proficiency in C and C++ for embedded software development and one or more modern programming languages like Golang, Python, Node, and Java. Hands-on experience with hardware-level security testing techniques, including JTAG/SWD, UART, SPI/I2C debug interfaces, firmware extraction, fault injection, and side-channel analysis. Working knowledge of reverse engineering and analysis tooling such as Ghidra, IDA Pro, Binary Ninja, binwalk, QEMU, and common debuggers across ARM, MIPS, and RISC-V architectures. At Optimum, every action and interaction we take part in, is driven by our three Guiding Principles: Do What's Right, Drive One Optimum, and Make It Happen. These aren't just words, they help us build trust, create real community, and embrace new ways of thinking. Our employees are empowered to do the right thing for our customers and co-workers and to recognize and reward these behaviors when we see them. It's all part of the bigger picture of "Be The Difference" where each employee knows they have the power to enact real change, share new ideas, and understand that learning never stop. If you have the drive to succeed and are ready to embark on a thrilling career, seize this opportunity today, and join our winning team. Together, we'll shape the future of connectivity. All job descriptions and required skills, qualifications and responsibilities for a particular position are subject to modification by the Company from time to time, in the Company's discretion based on business necessity. We are an Equal Opportunity Employer committed to recruiting, hiring and promoting qualified people of all backgrounds regardless of gender, race, color, creed, national origin, religion, age, marital status, pregnancy, physical or mental disability, sexual orientation, gender identity, military or veteran status, or any other basis protected by federal, state, or local law. The Company collects personal information about its applicants for employment that may include personal identifiers, professional or employment related information, photos, education information and/or protected classifications under federal and state law. This information is collected for employment purposes, including identification, work authorization, FCRA-compliant background screening, human resource administration and compliance with federal, state and local law. Applicants for employment with The Company will never be asked to provide money (even if reimbursable) as part of the job application or hiring process. Please review our Fraud FAQ for further details. We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.
09/23/2026
Full time
Are you looking to Optimize your life? Start your exciting path to a rewarding career today! We are Optimum, a leader in the fast-paced world of connectivity, and we're seeking driven and enthusiastic professionals to join our team, empower lives, fuel businesses, and drive innovation. Connectivity is now longer a luxury, but a necessity. A career at Optimum means you'll be enabling progress and enhancing lives by providing reliable, high-speed connectivity solutions that keep the world connected. Our successes, now and in the future, are powered by our amazing product, a commitment to our people and culture, and the connections we make in our communities. If you are resourceful, collaborative, and passionate about delivering consistent excellence, Optimum is for you! Job Summary Optimum is a leading provider of Mobile, Broadband (DOCSIS, Fiber) and Video services in the United States for Business and Residential Customers. We are looking for a passionate and engaged individual who is willing to take on the challenge of scaling product security at a Fortune 500 company and building the cyber security foundation that will allow our development teams to Go Fast! The Product Security organization helps Optimum move faster, securely. We're a team of engineers who work to enable other teams to build products as quickly as possible while continuing to protect our customers. We support developers in shipping secure code by building security tools and services, providing security training and expertise, and advocating for best practices in authentication, authorization, and safe data handling across the company. As a Product Security Engineer focusing on embedded systems security for video, broadband, and Wi-Fi products, you'll be a trusted partner, collaborating closely with engineering and product teams to ensure security is a cornerstone of every product. You will partner with leadership to shape product strategy, advocate for strong security controls, and influence future product iterations, and serve as an industry expert in device security engineering practices and standards. By leveraging your deep industry knowledge, you'll lead the charge in implementing secure architecture and design principles, ensuring early detection and prevention of vulnerabilities. Your expertise in security assessments and software engineering will help identify and mitigate potential threats, while your mentorship and training efforts will foster a security-first culture. Responsibilities Collaborate with engineering and product teams to integrate security and secure-by-default guardrails into the product lifecycle, ensuring that security is a core consideration in all design and development decisions. Conduct Threat Modeling and Risk Assessments from the early stages of the product development lifecycle to identify, assess, and prioritize security risks, enabling proactive mitigation strategies. Perform rigorous security testing and reviews to uncover and address security weaknesses. Lead initiatives automating security processes from the developer workstation to cloud, SaaS, and datacenter environments. Foster a security-first culture by educating and empowering engineering and product teams through training, awareness campaigns, and mentorship, cultivating a strong security mindset. Stay updated on the latest security threats, vulnerabilities, and technology trends, and proactively implement improvements. Contribute to incident response efforts, investigate root causes, and implement corrective actions to minimize impact and prevent future occurrences. Conduct penetration testing, reverse engineering, and vulnerability research against firmware, hardware interfaces, and device software to surface real-world attack paths before adversaries do. Establish and operate Software Bill of Materials (SBOM) management, CVE triage, and coordinated vulnerability disclosure processes for connected devices throughout their lifecycle. Define and uphold secure-by-design requirements aligned to recognized IoT and embedded security standards (e.g., NIST IR 8259, ETSI EN , OWASP IoT Top 10). Qualifications Bachelor's degree in Computer Science, Electrical Engineering, or a related field. Master's degree is a plus. 5+ years of hands-on experience in software engineering and designing and delivering security-critical systems for internet-connected embedded devices. Proven expertise in embedded systems and product security, with a strong understanding of modern software development processes and methods, security best practices, threat modeling, and risk assessment. Excellent communication skills, both written and verbal, and the ability to communicate complex security concepts to technical and non-technical audiences, including senior leadership. Proven ability to establish credibility and build trust with engineers and operational staff. Expertise in conducting comprehensive threat modeling, risk assessments, and code reviews to identify and mitigate vulnerabilities. Experience utilizing and securing AI/ML models and AI-integrated solutions, a general understanding of AI concepts, AI governance and risk management, and a willingness to learn more. Proficiency in secure SDLC practices and practical experience with CI/CD pipelines and DevOps tools. Experience overseeing vulnerability and threat management at the platform and device levels. Strong understanding of cryptography and key management use cases. In-depth knowledge of networking protocols, peripheral and firmware security, secure boot, embedded Linux security, Android or iOS security, and PKI. Experience working with special purpose security hardware such as Trusted Platform Modules (TPMs) and Hardware Security Modules (HSMs). Proficiency in C and C++ for embedded software development and one or more modern programming languages like Golang, Python, Node, and Java. Hands-on experience with hardware-level security testing techniques, including JTAG/SWD, UART, SPI/I2C debug interfaces, firmware extraction, fault injection, and side-channel analysis. Working knowledge of reverse engineering and analysis tooling such as Ghidra, IDA Pro, Binary Ninja, binwalk, QEMU, and common debuggers across ARM, MIPS, and RISC-V architectures. At Optimum, every action and interaction we take part in, is driven by our three Guiding Principles: Do What's Right, Drive One Optimum, and Make It Happen. These aren't just words, they help us build trust, create real community, and embrace new ways of thinking. Our employees are empowered to do the right thing for our customers and co-workers and to recognize and reward these behaviors when we see them. It's all part of the bigger picture of "Be The Difference" where each employee knows they have the power to enact real change, share new ideas, and understand that learning never stop. If you have the drive to succeed and are ready to embark on a thrilling career, seize this opportunity today, and join our winning team. Together, we'll shape the future of connectivity. All job descriptions and required skills, qualifications and responsibilities for a particular position are subject to modification by the Company from time to time, in the Company's discretion based on business necessity. We are an Equal Opportunity Employer committed to recruiting, hiring and promoting qualified people of all backgrounds regardless of gender, race, color, creed, national origin, religion, age, marital status, pregnancy, physical or mental disability, sexual orientation, gender identity, military or veteran status, or any other basis protected by federal, state, or local law. The Company collects personal information about its applicants for employment that may include personal identifiers, professional or employment related information, photos, education information and/or protected classifications under federal and state law. This information is collected for employment purposes, including identification, work authorization, FCRA-compliant background screening, human resource administration and compliance with federal, state and local law. Applicants for employment with The Company will never be asked to provide money (even if reimbursable) as part of the job application or hiring process. Please review our Fraud FAQ for further details. We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.
Senior HPC/GPU Systems Engineer
Nscale San Francisco, California
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack - energy, data centres, GPU superclusters, orchestration, and AI services - delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world. At Nscale, our Support and Operations team plays a critical role in maintaining service availability, driving service reliability, and delivering rapid response to customer issues. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future. About the Role (Job Purpose) Senior Infrastructure Support Engineers are the senior technical escalation point within Infrastructure Support, owning the health of Nscale's GPU fleets and the high-performance fabrics that connect them. This is a hands-on L2/L3 role operating at the intersection of GPU hardware, east-west networking, Linux, and data centre operations - acting as the operational bridge between Support, DC Operations, and Engineering. You will: Own complex, ambiguous problems end-to-end and make decisive calls in a results-driven environment, taking calculated risks where speed matters. Communicate technical detail clearly, specifically, and concisely - to engineers, to customers, and to leadership. We treat communication quality as a core engineering skill, not a soft skill. Influence without authority and build strong relationships with senior stakeholders across the business to get things done. Grasp new technical concepts quickly, stay curious, and know which questions to ask to get up to speed fast. Bring discipline and organisation: evidence-led investigations, accurate records, clean handovers. Experience required: 6+ years in infrastructure, operations, or support engineering roles in production environments, including 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates. What You'll be Doing (Responsibilities) Join the Support duty rotation as a senior escalation point, collaborating with Infrastructure Engineering, CNPRE, Network Operations, and Product Engineering on incidents, investigations, and changes. Diagnose and remediate GPU node faults across the full stack - driver, firmware, and hardware layers - from nvidia-smi/DCGM and XID/RAS analysis through BMC/Redfish and out-of-band management to physical fault isolation and vendor RMA. Own east-west fabric health: run link-level diagnostics (mlxlink, ibdiagnet, or equivalent), isolate transceiver, optics, cabling, and switch-port faults, and validate topology across InfiniBand and RoCE/high-speed Ethernet fabrics. Investigate data-path issues on high-performance storage platforms (e.g. VAST), including storage-network interactions across clients, mounts, VIPs, and routing. Run structured, hypothesis-driven investigations; conduct root cause analysis for major incidents and drive long-term fixes to completion. Author and execute changes in live customer environments with proper risk assessment, peer review, and backout plans. Proactively improve dashboards, alerts, and runbooks to prevent repeat incidents; identify recurring patterns and convert them into problem records and automation. Accurately record, update, and resolve tickets, keeping internal and external parties informed with clear customer-impact statements and evidence-rich notes that enable clean handover. Design and implement automation scripts and small tools to reduce toil and human intervention. Act as a key escalation point for the Support Organisation, taking ownership of strategic decisions where results matter. Mentor and upskill mid-level engineers; contribute to knowledge sharing across Operations and Engineering, including training content, workshops, and PR reviews. Lead by earning trust and speaking candidly. Disagree when appropriate and challenge the status quo; commit wholly to decisions once in motion. Respond to critical incidents out of business hours and participate in on-call as required. Travel to Nscale or customer sites to provide onsite technical expertise. About You (Skills / Qualifications Experience) Experience. 6+ years in infrastructure, operations, or support engineering in production environments; 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates, ideally in a customer-facing or escalation-driven capacity. Communication. Able to explain complex technical detail clearly, specifically, and concisely - in tickets, in incident updates, and face to face with customers and stakeholders at all levels. Strong written discipline: your notes let the next engineer pick up where you left off without starting from scratch. GPU platforms (NVIDIA; AMD Instinct beneficial). Practical, current experience with GPU drivers, firmware, and runtime stacks on AI training and inference clusters. Confident with nvidia-smi, DCGM, and XID/error interpretation; able to isolate faults across GPU, baseboard, NIC, and PCIe layers and drive them through diagnosis to RMA. High-performance east-west fabrics. Hands-on experience with RDMA fabrics - InfiniBand and/or RoCE - including link-layer diagnostics (mlxlink, ibdiagnet, or equivalent), transceiver and cabling fault isolation, and understanding of rail-optimised topologies, NVLink/NVSwitch, and NCCL-based performance troubleshooting on multi-node clusters. HPC scheduling. Slurm operations for large multi-GPU jobs - containers via Pyxis/Enroot, MPI, and diagnosing queue, topology, and job failures. Linux systems engineering at scale. Strong command of modern Linux distributions, kernel modules, systemd, networking stack, and filesystem tooling. Proven troubleshooting across compute, storage, and network layers in production. Server hardware and control planes. Comfortable with BMC/Redfish, firmware management, and bare-metal provisioning workflows (MAAS or similar) across large node fleets. Networking fundamentals. Solid grasp of L2/L3, routing, BGP, VLANs, VXLAN, firewalls, and load balancing, with a clear understanding of how east-west cluster traffic differs from north-south. Observability and incident response. Build and use alerting stacks and dashboards (Prometheus/Grafana or similar), interpret metrics and alerts, drive runbooks to resolution, and contribute to SLOs and post-incident reviews. Change and risk judgment. Experience authoring and executing changes in business-critical environments, including risk assessments, customer-impact analysis, and backout plans. SRE-style operations. Write and maintain runbooks, automate diagnostics, and reduce human intervention through scripts and small tools. Automation and Git. Scripting skills in Bash, Python, or equivalent for operational tooling and integrations; experience with infrastructure automation tools (Ansible, Terraform, or similar). Data centre fundamentals. Understanding of how data centres operate - servers, networks, storage, power, and cooling - ideally gained through an operational support background. Leadership. Disciplined, organised, and self-motivated, with the ability to mentor and motivate other engineers, take decisive action, and drive the team and wider organisation to improve. Adaptability. Able to adapt to customer-driven demands, including specialist support outside core hours and travel for onsite work. Nice to Have High-performance storage. Hands-on experience with VAST or comparable AI-optimised storage platforms, or Ceph/parallel filesystems and NFS at scale (multipath, remoteports, nconnect), including diagnosing storage-network interaction and data-path performance issues. OpenStack and fleet operations tooling. OpenStack operations experience (Neutron, Cinder, error triage), plus familiarity with fleet-scale tooling for provisioning, health, and remediation across large GPU estates (MAAS, NetBox, Redfish-driven automation, or similar). Kubernetes. Operating and troubleshooting clusters, including GPU operator stacks and understanding how physical resources are abstracted up the stack. Helpful context for our platform, though not the core of this role. Automation at scale. Automated network configuration with safe, repeatable changes in business-critical environments; GitOps and CI/CD pipelines (GitHub Actions or similar); access and security tooling such as Teleport or Vault in production. Certifications. Relevant GPU/HPC, datacenter architecture, Linux, networking, Kubernetes, cloud, or security certifications (e.g. RHCSA/RHCE, CKA, NVIDIA-certified) are a plus. What We Can Offer You At Nscale, you'll find a collaborative, supportive . click apply for full job details
09/23/2026
Full time
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack - energy, data centres, GPU superclusters, orchestration, and AI services - delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world. At Nscale, our Support and Operations team plays a critical role in maintaining service availability, driving service reliability, and delivering rapid response to customer issues. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future. About the Role (Job Purpose) Senior Infrastructure Support Engineers are the senior technical escalation point within Infrastructure Support, owning the health of Nscale's GPU fleets and the high-performance fabrics that connect them. This is a hands-on L2/L3 role operating at the intersection of GPU hardware, east-west networking, Linux, and data centre operations - acting as the operational bridge between Support, DC Operations, and Engineering. You will: Own complex, ambiguous problems end-to-end and make decisive calls in a results-driven environment, taking calculated risks where speed matters. Communicate technical detail clearly, specifically, and concisely - to engineers, to customers, and to leadership. We treat communication quality as a core engineering skill, not a soft skill. Influence without authority and build strong relationships with senior stakeholders across the business to get things done. Grasp new technical concepts quickly, stay curious, and know which questions to ask to get up to speed fast. Bring discipline and organisation: evidence-led investigations, accurate records, clean handovers. Experience required: 6+ years in infrastructure, operations, or support engineering roles in production environments, including 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates. What You'll be Doing (Responsibilities) Join the Support duty rotation as a senior escalation point, collaborating with Infrastructure Engineering, CNPRE, Network Operations, and Product Engineering on incidents, investigations, and changes. Diagnose and remediate GPU node faults across the full stack - driver, firmware, and hardware layers - from nvidia-smi/DCGM and XID/RAS analysis through BMC/Redfish and out-of-band management to physical fault isolation and vendor RMA. Own east-west fabric health: run link-level diagnostics (mlxlink, ibdiagnet, or equivalent), isolate transceiver, optics, cabling, and switch-port faults, and validate topology across InfiniBand and RoCE/high-speed Ethernet fabrics. Investigate data-path issues on high-performance storage platforms (e.g. VAST), including storage-network interactions across clients, mounts, VIPs, and routing. Run structured, hypothesis-driven investigations; conduct root cause analysis for major incidents and drive long-term fixes to completion. Author and execute changes in live customer environments with proper risk assessment, peer review, and backout plans. Proactively improve dashboards, alerts, and runbooks to prevent repeat incidents; identify recurring patterns and convert them into problem records and automation. Accurately record, update, and resolve tickets, keeping internal and external parties informed with clear customer-impact statements and evidence-rich notes that enable clean handover. Design and implement automation scripts and small tools to reduce toil and human intervention. Act as a key escalation point for the Support Organisation, taking ownership of strategic decisions where results matter. Mentor and upskill mid-level engineers; contribute to knowledge sharing across Operations and Engineering, including training content, workshops, and PR reviews. Lead by earning trust and speaking candidly. Disagree when appropriate and challenge the status quo; commit wholly to decisions once in motion. Respond to critical incidents out of business hours and participate in on-call as required. Travel to Nscale or customer sites to provide onsite technical expertise. About You (Skills / Qualifications Experience) Experience. 6+ years in infrastructure, operations, or support engineering in production environments; 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates, ideally in a customer-facing or escalation-driven capacity. Communication. Able to explain complex technical detail clearly, specifically, and concisely - in tickets, in incident updates, and face to face with customers and stakeholders at all levels. Strong written discipline: your notes let the next engineer pick up where you left off without starting from scratch. GPU platforms (NVIDIA; AMD Instinct beneficial). Practical, current experience with GPU drivers, firmware, and runtime stacks on AI training and inference clusters. Confident with nvidia-smi, DCGM, and XID/error interpretation; able to isolate faults across GPU, baseboard, NIC, and PCIe layers and drive them through diagnosis to RMA. High-performance east-west fabrics. Hands-on experience with RDMA fabrics - InfiniBand and/or RoCE - including link-layer diagnostics (mlxlink, ibdiagnet, or equivalent), transceiver and cabling fault isolation, and understanding of rail-optimised topologies, NVLink/NVSwitch, and NCCL-based performance troubleshooting on multi-node clusters. HPC scheduling. Slurm operations for large multi-GPU jobs - containers via Pyxis/Enroot, MPI, and diagnosing queue, topology, and job failures. Linux systems engineering at scale. Strong command of modern Linux distributions, kernel modules, systemd, networking stack, and filesystem tooling. Proven troubleshooting across compute, storage, and network layers in production. Server hardware and control planes. Comfortable with BMC/Redfish, firmware management, and bare-metal provisioning workflows (MAAS or similar) across large node fleets. Networking fundamentals. Solid grasp of L2/L3, routing, BGP, VLANs, VXLAN, firewalls, and load balancing, with a clear understanding of how east-west cluster traffic differs from north-south. Observability and incident response. Build and use alerting stacks and dashboards (Prometheus/Grafana or similar), interpret metrics and alerts, drive runbooks to resolution, and contribute to SLOs and post-incident reviews. Change and risk judgment. Experience authoring and executing changes in business-critical environments, including risk assessments, customer-impact analysis, and backout plans. SRE-style operations. Write and maintain runbooks, automate diagnostics, and reduce human intervention through scripts and small tools. Automation and Git. Scripting skills in Bash, Python, or equivalent for operational tooling and integrations; experience with infrastructure automation tools (Ansible, Terraform, or similar). Data centre fundamentals. Understanding of how data centres operate - servers, networks, storage, power, and cooling - ideally gained through an operational support background. Leadership. Disciplined, organised, and self-motivated, with the ability to mentor and motivate other engineers, take decisive action, and drive the team and wider organisation to improve. Adaptability. Able to adapt to customer-driven demands, including specialist support outside core hours and travel for onsite work. Nice to Have High-performance storage. Hands-on experience with VAST or comparable AI-optimised storage platforms, or Ceph/parallel filesystems and NFS at scale (multipath, remoteports, nconnect), including diagnosing storage-network interaction and data-path performance issues. OpenStack and fleet operations tooling. OpenStack operations experience (Neutron, Cinder, error triage), plus familiarity with fleet-scale tooling for provisioning, health, and remediation across large GPU estates (MAAS, NetBox, Redfish-driven automation, or similar). Kubernetes. Operating and troubleshooting clusters, including GPU operator stacks and understanding how physical resources are abstracted up the stack. Helpful context for our platform, though not the core of this role. Automation at scale. Automated network configuration with safe, repeatable changes in business-critical environments; GitOps and CI/CD pipelines (GitHub Actions or similar); access and security tooling such as Teleport or Vault in production. Certifications. Relevant GPU/HPC, datacenter architecture, Linux, networking, Kubernetes, cloud, or security certifications (e.g. RHCSA/RHCE, CKA, NVIDIA-certified) are a plus. What We Can Offer You At Nscale, you'll find a collaborative, supportive . click apply for full job details
Senior Technical Product Manager, GPU Infrastructure
Nscale New York, New York
About Nscale Nscale is taking on the hyperscalers by building a vertically integrated GenAI cloud platform. We own the data centres, software, and applications that power today's AI stack using sustainable technology solutions. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. Collaboration is key, and we work together swiftly and respectfully, embracing adaptability and resilience in all we do. About the role Technical Product Managers at Nscale own the definition, delivery, and ongoing evolution of a slice of the Nscale platform. You partner closely with engineering, design, research, and go-to-market teams to translate customer problems and operational realities into shippable product outcomes. As a Senior Technical Product Manager for Fleet Operations, you own the product strategy for the day 0-2+ operational software that runs our global GPU fleet - the systems that bring capacity online, keep it healthy, and restore it fast when things go wrong. You partner daily with Fleet Software engineering teams, SRE, and Support to turn operational pain into durable product: provisioning and bringup (day 0), testing and deployment (day 1), and the full lifecycle of monitoring, incident response, repair, RMA, firmware, and decommissioning (day 2+). You operate at team scope, owning a major product area and driving multi-quarter initiatives that directly move fleet availability, utilisation, and time-to-recover. Senior Technical Product Manager, Fleet Operations 1 What you'll be doing Own the strategy and roadmap for a significant Fleet Operations product area - e.g. provisioning and bring-up, fleet health and telemetry, incident and repair workflows, firmware and lifecycle management, or capacity and inventory. Lead multi-sprint, cross-functional initiatives from problem framing through rollout across live GPU clusters, working hand-in-hand with Fleet Software, SRE, data centre operations, and Support. Turn operational ambiguity into product: shadow on-call rotations, ride along with support and repair workflows, and translate recurring toil into tooling, automation, and platform capabilities. Define the metrics that matter for a GPU fleet - availability, utilisation, MTTR, time-to-bring-up, hardware failure rates, support ticket deflection - and drive the roadmap against them. Partner with engineering on architecture and trade-offs for systems that span bare metal, orchestration, observability, and control planes. Drive incident reviews and postmortems into product commitments; close the loop so the same class of issue doesn't recur. Mentor junior product managers and raise the quality bar for PRDs, reviews, and product decisions across the team. Represent Fleet Operations in planning, reviews, and leadership updates. What you need 5-8 years of product management experience in software or technology, with a track record of owning significant product areas in infrastructure, platform, or operations-facing products. Strong technical fluency in large-scale systems: you can lead discussions with engineering on architecture, trade-offs, and feasibility across provisioning, orchestration, observability, and control-plane design Experience building products for operators - SREs, NOC/support teams, data centre technicians, or similar - and a genuine appetite for understanding their workflows. Demonstrated ability to move from an ambiguous operational problem space to shipped product outcomes that measurably improve reliability, efficiency, or time-to-recover. Experience mentoring or informally leading peers. Excellent written and verbal communication; you can make complex product decisions legible to engineers, operators, and executives alike. Experience with data centre networking technologies, including high-performance GPU interconnects such as InfiniBand and RoCE (RDMA over Converged Ethernet), and an understanding of how backend (east-west/compute) and frontend (north-south/storage and management) network fabrics are designed and operated at scale. Familiarity with WAN, edge, and global backbone architectures - including how multi-site connectivity, peering, and traffic engineering support a globally distributed GPU fleet. Experience partnering with network engineering teams on fabric health, congestion monitoring, and link-level failure workflows, ideally in environments where network performance directly impacts training or inference workloads. Nice to haves Degree in computer science, engineering, or a related field, or prior experience as an engineer or SRE. Hands-on background in cloud infrastructure, bare-metal provisioning, fleet or hardware lifecycle management, observability/monitoring platforms, or incident management tooling. Experience with bare-metal provisioning systems such as OpenStack Ironic (or equivalents like MAAS, Tinkerbell, or in-house provisioning stacks). Experience with DCIM tools such as NetBox (or equivalents like Device42 or Nautobot) for inventory, cabling, and rack/asset management. Experience with ITSM and ticketing platforms such as Jira Service Management (or equivalents like ServiceNow, Zendesk, or Freshservice) for support, incident, and RMA workflows. Experience with observability and monitoring platforms such as Grafana, Prometheus, Datadog, or equivalents - ideally including defining SLOs, dashboards, and alerting for large fleets. Familiarity with GPU or accelerated compute environments, data centre operations, or hyperscaler-style fleet management. Experience operating in high-growth or early-stage environments where the product is being built alongside the fleet itself Join Nscale as we build a world-class AI cloud platform. If you're excited about owning the software that keeps a global GPU fleet running - and raising the bar for the team around you - we'd love to hear from you! At Nscale, we are committed to fostering an inclusive, diverse, and equitable workplace. We believe that a variety of perspectives enriches our work environment, and we encourage applications from candidates of all backgrounds, experiences, and abilities. We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds. If there's anything we can do to accommodate your specific situation, please let us know. The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role. The range below reflects the base salary for the position. Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs. Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation. Salary Range $200,000-$280,000 USD For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here. Nscale does not accept unsolicited candidate submissions from recruitment agencies.
09/23/2026
Full time
About Nscale Nscale is taking on the hyperscalers by building a vertically integrated GenAI cloud platform. We own the data centres, software, and applications that power today's AI stack using sustainable technology solutions. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. Collaboration is key, and we work together swiftly and respectfully, embracing adaptability and resilience in all we do. About the role Technical Product Managers at Nscale own the definition, delivery, and ongoing evolution of a slice of the Nscale platform. You partner closely with engineering, design, research, and go-to-market teams to translate customer problems and operational realities into shippable product outcomes. As a Senior Technical Product Manager for Fleet Operations, you own the product strategy for the day 0-2+ operational software that runs our global GPU fleet - the systems that bring capacity online, keep it healthy, and restore it fast when things go wrong. You partner daily with Fleet Software engineering teams, SRE, and Support to turn operational pain into durable product: provisioning and bringup (day 0), testing and deployment (day 1), and the full lifecycle of monitoring, incident response, repair, RMA, firmware, and decommissioning (day 2+). You operate at team scope, owning a major product area and driving multi-quarter initiatives that directly move fleet availability, utilisation, and time-to-recover. Senior Technical Product Manager, Fleet Operations 1 What you'll be doing Own the strategy and roadmap for a significant Fleet Operations product area - e.g. provisioning and bring-up, fleet health and telemetry, incident and repair workflows, firmware and lifecycle management, or capacity and inventory. Lead multi-sprint, cross-functional initiatives from problem framing through rollout across live GPU clusters, working hand-in-hand with Fleet Software, SRE, data centre operations, and Support. Turn operational ambiguity into product: shadow on-call rotations, ride along with support and repair workflows, and translate recurring toil into tooling, automation, and platform capabilities. Define the metrics that matter for a GPU fleet - availability, utilisation, MTTR, time-to-bring-up, hardware failure rates, support ticket deflection - and drive the roadmap against them. Partner with engineering on architecture and trade-offs for systems that span bare metal, orchestration, observability, and control planes. Drive incident reviews and postmortems into product commitments; close the loop so the same class of issue doesn't recur. Mentor junior product managers and raise the quality bar for PRDs, reviews, and product decisions across the team. Represent Fleet Operations in planning, reviews, and leadership updates. What you need 5-8 years of product management experience in software or technology, with a track record of owning significant product areas in infrastructure, platform, or operations-facing products. Strong technical fluency in large-scale systems: you can lead discussions with engineering on architecture, trade-offs, and feasibility across provisioning, orchestration, observability, and control-plane design Experience building products for operators - SREs, NOC/support teams, data centre technicians, or similar - and a genuine appetite for understanding their workflows. Demonstrated ability to move from an ambiguous operational problem space to shipped product outcomes that measurably improve reliability, efficiency, or time-to-recover. Experience mentoring or informally leading peers. Excellent written and verbal communication; you can make complex product decisions legible to engineers, operators, and executives alike. Experience with data centre networking technologies, including high-performance GPU interconnects such as InfiniBand and RoCE (RDMA over Converged Ethernet), and an understanding of how backend (east-west/compute) and frontend (north-south/storage and management) network fabrics are designed and operated at scale. Familiarity with WAN, edge, and global backbone architectures - including how multi-site connectivity, peering, and traffic engineering support a globally distributed GPU fleet. Experience partnering with network engineering teams on fabric health, congestion monitoring, and link-level failure workflows, ideally in environments where network performance directly impacts training or inference workloads. Nice to haves Degree in computer science, engineering, or a related field, or prior experience as an engineer or SRE. Hands-on background in cloud infrastructure, bare-metal provisioning, fleet or hardware lifecycle management, observability/monitoring platforms, or incident management tooling. Experience with bare-metal provisioning systems such as OpenStack Ironic (or equivalents like MAAS, Tinkerbell, or in-house provisioning stacks). Experience with DCIM tools such as NetBox (or equivalents like Device42 or Nautobot) for inventory, cabling, and rack/asset management. Experience with ITSM and ticketing platforms such as Jira Service Management (or equivalents like ServiceNow, Zendesk, or Freshservice) for support, incident, and RMA workflows. Experience with observability and monitoring platforms such as Grafana, Prometheus, Datadog, or equivalents - ideally including defining SLOs, dashboards, and alerting for large fleets. Familiarity with GPU or accelerated compute environments, data centre operations, or hyperscaler-style fleet management. Experience operating in high-growth or early-stage environments where the product is being built alongside the fleet itself Join Nscale as we build a world-class AI cloud platform. If you're excited about owning the software that keeps a global GPU fleet running - and raising the bar for the team around you - we'd love to hear from you! At Nscale, we are committed to fostering an inclusive, diverse, and equitable workplace. We believe that a variety of perspectives enriches our work environment, and we encourage applications from candidates of all backgrounds, experiences, and abilities. We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds. If there's anything we can do to accommodate your specific situation, please let us know. The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role. The range below reflects the base salary for the position. Actual compensation may vary based on job-related factors such as skill set, experience, education, and location. In addition to base salary, this role may be eligible for bonus, equity, and/or commission programs. Nscale may offer a competitive benefits package including medical, dental, vision, flexible paid time off, parental leave, and retirement plan participation. Salary Range $200,000-$280,000 USD For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here. Nscale does not accept unsolicited candidate submissions from recruitment agencies.
Senior HPC/GPU Systems Engineer
Nscale Seattle, Washington
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack - energy, data centres, GPU superclusters, orchestration, and AI services - delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world. At Nscale, our Support and Operations team plays a critical role in maintaining service availability, driving service reliability, and delivering rapid response to customer issues. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future. About the Role (Job Purpose) Senior Infrastructure Support Engineers are the senior technical escalation point within Infrastructure Support, owning the health of Nscale's GPU fleets and the high-performance fabrics that connect them. This is a hands-on L2/L3 role operating at the intersection of GPU hardware, east-west networking, Linux, and data centre operations - acting as the operational bridge between Support, DC Operations, and Engineering. You will: Own complex, ambiguous problems end-to-end and make decisive calls in a results-driven environment, taking calculated risks where speed matters. Communicate technical detail clearly, specifically, and concisely - to engineers, to customers, and to leadership. We treat communication quality as a core engineering skill, not a soft skill. Influence without authority and build strong relationships with senior stakeholders across the business to get things done. Grasp new technical concepts quickly, stay curious, and know which questions to ask to get up to speed fast. Bring discipline and organisation: evidence-led investigations, accurate records, clean handovers. Experience required: 6+ years in infrastructure, operations, or support engineering roles in production environments, including 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates. What You'll be Doing (Responsibilities) Join the Support duty rotation as a senior escalation point, collaborating with Infrastructure Engineering, CNPRE, Network Operations, and Product Engineering on incidents, investigations, and changes. Diagnose and remediate GPU node faults across the full stack - driver, firmware, and hardware layers - from nvidia-smi/DCGM and XID/RAS analysis through BMC/Redfish and out-of-band management to physical fault isolation and vendor RMA. Own east-west fabric health: run link-level diagnostics (mlxlink, ibdiagnet, or equivalent), isolate transceiver, optics, cabling, and switch-port faults, and validate topology across InfiniBand and RoCE/high-speed Ethernet fabrics. Investigate data-path issues on high-performance storage platforms (e.g. VAST), including storage-network interactions across clients, mounts, VIPs, and routing. Run structured, hypothesis-driven investigations; conduct root cause analysis for major incidents and drive long-term fixes to completion. Author and execute changes in live customer environments with proper risk assessment, peer review, and backout plans. Proactively improve dashboards, alerts, and runbooks to prevent repeat incidents; identify recurring patterns and convert them into problem records and automation. Accurately record, update, and resolve tickets, keeping internal and external parties informed with clear customer-impact statements and evidence-rich notes that enable clean handover. Design and implement automation scripts and small tools to reduce toil and human intervention. Act as a key escalation point for the Support Organisation, taking ownership of strategic decisions where results matter. Mentor and upskill mid-level engineers; contribute to knowledge sharing across Operations and Engineering, including training content, workshops, and PR reviews. Lead by earning trust and speaking candidly. Disagree when appropriate and challenge the status quo; commit wholly to decisions once in motion. Respond to critical incidents out of business hours and participate in on-call as required. Travel to Nscale or customer sites to provide onsite technical expertise. About You (Skills / Qualifications Experience) Experience. 6+ years in infrastructure, operations, or support engineering in production environments; 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates, ideally in a customer-facing or escalation-driven capacity. Communication. Able to explain complex technical detail clearly, specifically, and concisely - in tickets, in incident updates, and face to face with customers and stakeholders at all levels. Strong written discipline: your notes let the next engineer pick up where you left off without starting from scratch. GPU platforms (NVIDIA; AMD Instinct beneficial). Practical, current experience with GPU drivers, firmware, and runtime stacks on AI training and inference clusters. Confident with nvidia-smi, DCGM, and XID/error interpretation; able to isolate faults across GPU, baseboard, NIC, and PCIe layers and drive them through diagnosis to RMA. High-performance east-west fabrics. Hands-on experience with RDMA fabrics - InfiniBand and/or RoCE - including link-layer diagnostics (mlxlink, ibdiagnet, or equivalent), transceiver and cabling fault isolation, and understanding of rail-optimised topologies, NVLink/NVSwitch, and NCCL-based performance troubleshooting on multi-node clusters. HPC scheduling. Slurm operations for large multi-GPU jobs - containers via Pyxis/Enroot, MPI, and diagnosing queue, topology, and job failures. Linux systems engineering at scale. Strong command of modern Linux distributions, kernel modules, systemd, networking stack, and filesystem tooling. Proven troubleshooting across compute, storage, and network layers in production. Server hardware and control planes. Comfortable with BMC/Redfish, firmware management, and bare-metal provisioning workflows (MAAS or similar) across large node fleets. Networking fundamentals. Solid grasp of L2/L3, routing, BGP, VLANs, VXLAN, firewalls, and load balancing, with a clear understanding of how east-west cluster traffic differs from north-south. Observability and incident response. Build and use alerting stacks and dashboards (Prometheus/Grafana or similar), interpret metrics and alerts, drive runbooks to resolution, and contribute to SLOs and post-incident reviews. Change and risk judgment. Experience authoring and executing changes in business-critical environments, including risk assessments, customer-impact analysis, and backout plans. SRE-style operations. Write and maintain runbooks, automate diagnostics, and reduce human intervention through scripts and small tools. Automation and Git. Scripting skills in Bash, Python, or equivalent for operational tooling and integrations; experience with infrastructure automation tools (Ansible, Terraform, or similar). Data centre fundamentals. Understanding of how data centres operate - servers, networks, storage, power, and cooling - ideally gained through an operational support background. Leadership. Disciplined, organised, and self-motivated, with the ability to mentor and motivate other engineers, take decisive action, and drive the team and wider organisation to improve. Adaptability. Able to adapt to customer-driven demands, including specialist support outside core hours and travel for onsite work. Nice to Have High-performance storage. Hands-on experience with VAST or comparable AI-optimised storage platforms, or Ceph/parallel filesystems and NFS at scale (multipath, remoteports, nconnect), including diagnosing storage-network interaction and data-path performance issues. OpenStack and fleet operations tooling. OpenStack operations experience (Neutron, Cinder, error triage), plus familiarity with fleet-scale tooling for provisioning, health, and remediation across large GPU estates (MAAS, NetBox, Redfish-driven automation, or similar). Kubernetes. Operating and troubleshooting clusters, including GPU operator stacks and understanding how physical resources are abstracted up the stack. Helpful context for our platform, though not the core of this role. Automation at scale. Automated network configuration with safe, repeatable changes in business-critical environments; GitOps and CI/CD pipelines (GitHub Actions or similar); access and security tooling such as Teleport or Vault in production. Certifications. Relevant GPU/HPC, datacenter architecture, Linux, networking, Kubernetes, cloud, or security certifications (e.g. RHCSA/RHCE, CKA, NVIDIA-certified) are a plus. What We Can Offer You At Nscale, you'll find a collaborative, supportive . click apply for full job details
09/23/2026
Full time
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack - energy, data centres, GPU superclusters, orchestration, and AI services - delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world. At Nscale, our Support and Operations team plays a critical role in maintaining service availability, driving service reliability, and delivering rapid response to customer issues. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future. About the Role (Job Purpose) Senior Infrastructure Support Engineers are the senior technical escalation point within Infrastructure Support, owning the health of Nscale's GPU fleets and the high-performance fabrics that connect them. This is a hands-on L2/L3 role operating at the intersection of GPU hardware, east-west networking, Linux, and data centre operations - acting as the operational bridge between Support, DC Operations, and Engineering. You will: Own complex, ambiguous problems end-to-end and make decisive calls in a results-driven environment, taking calculated risks where speed matters. Communicate technical detail clearly, specifically, and concisely - to engineers, to customers, and to leadership. We treat communication quality as a core engineering skill, not a soft skill. Influence without authority and build strong relationships with senior stakeholders across the business to get things done. Grasp new technical concepts quickly, stay curious, and know which questions to ask to get up to speed fast. Bring discipline and organisation: evidence-led investigations, accurate records, clean handovers. Experience required: 6+ years in infrastructure, operations, or support engineering roles in production environments, including 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates. What You'll be Doing (Responsibilities) Join the Support duty rotation as a senior escalation point, collaborating with Infrastructure Engineering, CNPRE, Network Operations, and Product Engineering on incidents, investigations, and changes. Diagnose and remediate GPU node faults across the full stack - driver, firmware, and hardware layers - from nvidia-smi/DCGM and XID/RAS analysis through BMC/Redfish and out-of-band management to physical fault isolation and vendor RMA. Own east-west fabric health: run link-level diagnostics (mlxlink, ibdiagnet, or equivalent), isolate transceiver, optics, cabling, and switch-port faults, and validate topology across InfiniBand and RoCE/high-speed Ethernet fabrics. Investigate data-path issues on high-performance storage platforms (e.g. VAST), including storage-network interactions across clients, mounts, VIPs, and routing. Run structured, hypothesis-driven investigations; conduct root cause analysis for major incidents and drive long-term fixes to completion. Author and execute changes in live customer environments with proper risk assessment, peer review, and backout plans. Proactively improve dashboards, alerts, and runbooks to prevent repeat incidents; identify recurring patterns and convert them into problem records and automation. Accurately record, update, and resolve tickets, keeping internal and external parties informed with clear customer-impact statements and evidence-rich notes that enable clean handover. Design and implement automation scripts and small tools to reduce toil and human intervention. Act as a key escalation point for the Support Organisation, taking ownership of strategic decisions where results matter. Mentor and upskill mid-level engineers; contribute to knowledge sharing across Operations and Engineering, including training content, workshops, and PR reviews. Lead by earning trust and speaking candidly. Disagree when appropriate and challenge the status quo; commit wholly to decisions once in motion. Respond to critical incidents out of business hours and participate in on-call as required. Travel to Nscale or customer sites to provide onsite technical expertise. About You (Skills / Qualifications Experience) Experience. 6+ years in infrastructure, operations, or support engineering in production environments; 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates, ideally in a customer-facing or escalation-driven capacity. Communication. Able to explain complex technical detail clearly, specifically, and concisely - in tickets, in incident updates, and face to face with customers and stakeholders at all levels. Strong written discipline: your notes let the next engineer pick up where you left off without starting from scratch. GPU platforms (NVIDIA; AMD Instinct beneficial). Practical, current experience with GPU drivers, firmware, and runtime stacks on AI training and inference clusters. Confident with nvidia-smi, DCGM, and XID/error interpretation; able to isolate faults across GPU, baseboard, NIC, and PCIe layers and drive them through diagnosis to RMA. High-performance east-west fabrics. Hands-on experience with RDMA fabrics - InfiniBand and/or RoCE - including link-layer diagnostics (mlxlink, ibdiagnet, or equivalent), transceiver and cabling fault isolation, and understanding of rail-optimised topologies, NVLink/NVSwitch, and NCCL-based performance troubleshooting on multi-node clusters. HPC scheduling. Slurm operations for large multi-GPU jobs - containers via Pyxis/Enroot, MPI, and diagnosing queue, topology, and job failures. Linux systems engineering at scale. Strong command of modern Linux distributions, kernel modules, systemd, networking stack, and filesystem tooling. Proven troubleshooting across compute, storage, and network layers in production. Server hardware and control planes. Comfortable with BMC/Redfish, firmware management, and bare-metal provisioning workflows (MAAS or similar) across large node fleets. Networking fundamentals. Solid grasp of L2/L3, routing, BGP, VLANs, VXLAN, firewalls, and load balancing, with a clear understanding of how east-west cluster traffic differs from north-south. Observability and incident response. Build and use alerting stacks and dashboards (Prometheus/Grafana or similar), interpret metrics and alerts, drive runbooks to resolution, and contribute to SLOs and post-incident reviews. Change and risk judgment. Experience authoring and executing changes in business-critical environments, including risk assessments, customer-impact analysis, and backout plans. SRE-style operations. Write and maintain runbooks, automate diagnostics, and reduce human intervention through scripts and small tools. Automation and Git. Scripting skills in Bash, Python, or equivalent for operational tooling and integrations; experience with infrastructure automation tools (Ansible, Terraform, or similar). Data centre fundamentals. Understanding of how data centres operate - servers, networks, storage, power, and cooling - ideally gained through an operational support background. Leadership. Disciplined, organised, and self-motivated, with the ability to mentor and motivate other engineers, take decisive action, and drive the team and wider organisation to improve. Adaptability. Able to adapt to customer-driven demands, including specialist support outside core hours and travel for onsite work. Nice to Have High-performance storage. Hands-on experience with VAST or comparable AI-optimised storage platforms, or Ceph/parallel filesystems and NFS at scale (multipath, remoteports, nconnect), including diagnosing storage-network interaction and data-path performance issues. OpenStack and fleet operations tooling. OpenStack operations experience (Neutron, Cinder, error triage), plus familiarity with fleet-scale tooling for provisioning, health, and remediation across large GPU estates (MAAS, NetBox, Redfish-driven automation, or similar). Kubernetes. Operating and troubleshooting clusters, including GPU operator stacks and understanding how physical resources are abstracted up the stack. Helpful context for our platform, though not the core of this role. Automation at scale. Automated network configuration with safe, repeatable changes in business-critical environments; GitOps and CI/CD pipelines (GitHub Actions or similar); access and security tooling such as Teleport or Vault in production. Certifications. Relevant GPU/HPC, datacenter architecture, Linux, networking, Kubernetes, cloud, or security certifications (e.g. RHCSA/RHCE, CKA, NVIDIA-certified) are a plus. What We Can Offer You At Nscale, you'll find a collaborative, supportive . click apply for full job details
Senior HPC/GPU Systems Engineer
Nscale Houston, Texas
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack - energy, data centres, GPU superclusters, orchestration, and AI services - delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world. At Nscale, our Support and Operations team plays a critical role in maintaining service availability, driving service reliability, and delivering rapid response to customer issues. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future. About the Role (Job Purpose) Senior Infrastructure Support Engineers are the senior technical escalation point within Infrastructure Support, owning the health of Nscale's GPU fleets and the high-performance fabrics that connect them. This is a hands-on L2/L3 role operating at the intersection of GPU hardware, east-west networking, Linux, and data centre operations - acting as the operational bridge between Support, DC Operations, and Engineering. You will: Own complex, ambiguous problems end-to-end and make decisive calls in a results-driven environment, taking calculated risks where speed matters. Communicate technical detail clearly, specifically, and concisely - to engineers, to customers, and to leadership. We treat communication quality as a core engineering skill, not a soft skill. Influence without authority and build strong relationships with senior stakeholders across the business to get things done. Grasp new technical concepts quickly, stay curious, and know which questions to ask to get up to speed fast. Bring discipline and organisation: evidence-led investigations, accurate records, clean handovers. Experience required: 6+ years in infrastructure, operations, or support engineering roles in production environments, including 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates. What You'll be Doing (Responsibilities) Join the Support duty rotation as a senior escalation point, collaborating with Infrastructure Engineering, CNPRE, Network Operations, and Product Engineering on incidents, investigations, and changes. Diagnose and remediate GPU node faults across the full stack - driver, firmware, and hardware layers - from nvidia-smi/DCGM and XID/RAS analysis through BMC/Redfish and out-of-band management to physical fault isolation and vendor RMA. Own east-west fabric health: run link-level diagnostics (mlxlink, ibdiagnet, or equivalent), isolate transceiver, optics, cabling, and switch-port faults, and validate topology across InfiniBand and RoCE/high-speed Ethernet fabrics. Investigate data-path issues on high-performance storage platforms (e.g. VAST), including storage-network interactions across clients, mounts, VIPs, and routing. Run structured, hypothesis-driven investigations; conduct root cause analysis for major incidents and drive long-term fixes to completion. Author and execute changes in live customer environments with proper risk assessment, peer review, and backout plans. Proactively improve dashboards, alerts, and runbooks to prevent repeat incidents; identify recurring patterns and convert them into problem records and automation. Accurately record, update, and resolve tickets, keeping internal and external parties informed with clear customer-impact statements and evidence-rich notes that enable clean handover. Design and implement automation scripts and small tools to reduce toil and human intervention. Act as a key escalation point for the Support Organisation, taking ownership of strategic decisions where results matter. Mentor and upskill mid-level engineers; contribute to knowledge sharing across Operations and Engineering, including training content, workshops, and PR reviews. Lead by earning trust and speaking candidly. Disagree when appropriate and challenge the status quo; commit wholly to decisions once in motion. Respond to critical incidents out of business hours and participate in on-call as required. Travel to Nscale or customer sites to provide onsite technical expertise. About You (Skills / Qualifications Experience) Experience. 6+ years in infrastructure, operations, or support engineering in production environments; 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates, ideally in a customer-facing or escalation-driven capacity. Communication. Able to explain complex technical detail clearly, specifically, and concisely - in tickets, in incident updates, and face to face with customers and stakeholders at all levels. Strong written discipline: your notes let the next engineer pick up where you left off without starting from scratch. GPU platforms (NVIDIA; AMD Instinct beneficial). Practical, current experience with GPU drivers, firmware, and runtime stacks on AI training and inference clusters. Confident with nvidia-smi, DCGM, and XID/error interpretation; able to isolate faults across GPU, baseboard, NIC, and PCIe layers and drive them through diagnosis to RMA. High-performance east-west fabrics. Hands-on experience with RDMA fabrics - InfiniBand and/or RoCE - including link-layer diagnostics (mlxlink, ibdiagnet, or equivalent), transceiver and cabling fault isolation, and understanding of rail-optimised topologies, NVLink/NVSwitch, and NCCL-based performance troubleshooting on multi-node clusters. HPC scheduling. Slurm operations for large multi-GPU jobs - containers via Pyxis/Enroot, MPI, and diagnosing queue, topology, and job failures. Linux systems engineering at scale. Strong command of modern Linux distributions, kernel modules, systemd, networking stack, and filesystem tooling. Proven troubleshooting across compute, storage, and network layers in production. Server hardware and control planes. Comfortable with BMC/Redfish, firmware management, and bare-metal provisioning workflows (MAAS or similar) across large node fleets. Networking fundamentals. Solid grasp of L2/L3, routing, BGP, VLANs, VXLAN, firewalls, and load balancing, with a clear understanding of how east-west cluster traffic differs from north-south. Observability and incident response. Build and use alerting stacks and dashboards (Prometheus/Grafana or similar), interpret metrics and alerts, drive runbooks to resolution, and contribute to SLOs and post-incident reviews. Change and risk judgment. Experience authoring and executing changes in business-critical environments, including risk assessments, customer-impact analysis, and backout plans. SRE-style operations. Write and maintain runbooks, automate diagnostics, and reduce human intervention through scripts and small tools. Automation and Git. Scripting skills in Bash, Python, or equivalent for operational tooling and integrations; experience with infrastructure automation tools (Ansible, Terraform, or similar). Data centre fundamentals. Understanding of how data centres operate - servers, networks, storage, power, and cooling - ideally gained through an operational support background. Leadership. Disciplined, organised, and self-motivated, with the ability to mentor and motivate other engineers, take decisive action, and drive the team and wider organisation to improve. Adaptability. Able to adapt to customer-driven demands, including specialist support outside core hours and travel for onsite work. Nice to Have High-performance storage. Hands-on experience with VAST or comparable AI-optimised storage platforms, or Ceph/parallel filesystems and NFS at scale (multipath, remoteports, nconnect), including diagnosing storage-network interaction and data-path performance issues. OpenStack and fleet operations tooling. OpenStack operations experience (Neutron, Cinder, error triage), plus familiarity with fleet-scale tooling for provisioning, health, and remediation across large GPU estates (MAAS, NetBox, Redfish-driven automation, or similar). Kubernetes. Operating and troubleshooting clusters, including GPU operator stacks and understanding how physical resources are abstracted up the stack. Helpful context for our platform, though not the core of this role. Automation at scale. Automated network configuration with safe, repeatable changes in business-critical environments; GitOps and CI/CD pipelines (GitHub Actions or similar); access and security tooling such as Teleport or Vault in production. Certifications. Relevant GPU/HPC, datacenter architecture, Linux, networking, Kubernetes, cloud, or security certifications (e.g. RHCSA/RHCE, CKA, NVIDIA-certified) are a plus. What We Can Offer You At Nscale, you'll find a collaborative, supportive . click apply for full job details
09/23/2026
Full time
About Nscale Nscale is the vertically integrated AI cloud engineered for AI. We own and operate the full stack - energy, data centres, GPU superclusters, orchestration, and AI services - delivering high-performance infrastructure to AI-native companies, enterprises, and governments across Europe and the US. We are deploying GPU capacity at hyperscale, operating some of the densest, most advanced AI infrastructure in the world. At Nscale, our Support and Operations team plays a critical role in maintaining service availability, driving service reliability, and delivering rapid response to customer issues. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future. About the Role (Job Purpose) Senior Infrastructure Support Engineers are the senior technical escalation point within Infrastructure Support, owning the health of Nscale's GPU fleets and the high-performance fabrics that connect them. This is a hands-on L2/L3 role operating at the intersection of GPU hardware, east-west networking, Linux, and data centre operations - acting as the operational bridge between Support, DC Operations, and Engineering. You will: Own complex, ambiguous problems end-to-end and make decisive calls in a results-driven environment, taking calculated risks where speed matters. Communicate technical detail clearly, specifically, and concisely - to engineers, to customers, and to leadership. We treat communication quality as a core engineering skill, not a soft skill. Influence without authority and build strong relationships with senior stakeholders across the business to get things done. Grasp new technical concepts quickly, stay curious, and know which questions to ask to get up to speed fast. Bring discipline and organisation: evidence-led investigations, accurate records, clean handovers. Experience required: 6+ years in infrastructure, operations, or support engineering roles in production environments, including 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates. What You'll be Doing (Responsibilities) Join the Support duty rotation as a senior escalation point, collaborating with Infrastructure Engineering, CNPRE, Network Operations, and Product Engineering on incidents, investigations, and changes. Diagnose and remediate GPU node faults across the full stack - driver, firmware, and hardware layers - from nvidia-smi/DCGM and XID/RAS analysis through BMC/Redfish and out-of-band management to physical fault isolation and vendor RMA. Own east-west fabric health: run link-level diagnostics (mlxlink, ibdiagnet, or equivalent), isolate transceiver, optics, cabling, and switch-port faults, and validate topology across InfiniBand and RoCE/high-speed Ethernet fabrics. Investigate data-path issues on high-performance storage platforms (e.g. VAST), including storage-network interactions across clients, mounts, VIPs, and routing. Run structured, hypothesis-driven investigations; conduct root cause analysis for major incidents and drive long-term fixes to completion. Author and execute changes in live customer environments with proper risk assessment, peer review, and backout plans. Proactively improve dashboards, alerts, and runbooks to prevent repeat incidents; identify recurring patterns and convert them into problem records and automation. Accurately record, update, and resolve tickets, keeping internal and external parties informed with clear customer-impact statements and evidence-rich notes that enable clean handover. Design and implement automation scripts and small tools to reduce toil and human intervention. Act as a key escalation point for the Support Organisation, taking ownership of strategic decisions where results matter. Mentor and upskill mid-level engineers; contribute to knowledge sharing across Operations and Engineering, including training content, workshops, and PR reviews. Lead by earning trust and speaking candidly. Disagree when appropriate and challenge the status quo; commit wholly to decisions once in motion. Respond to critical incidents out of business hours and participate in on-call as required. Travel to Nscale or customer sites to provide onsite technical expertise. About You (Skills / Qualifications Experience) Experience. 6+ years in infrastructure, operations, or support engineering in production environments; 2-3+ years hands-on with GPU, HPC, or large-scale data centre estates, ideally in a customer-facing or escalation-driven capacity. Communication. Able to explain complex technical detail clearly, specifically, and concisely - in tickets, in incident updates, and face to face with customers and stakeholders at all levels. Strong written discipline: your notes let the next engineer pick up where you left off without starting from scratch. GPU platforms (NVIDIA; AMD Instinct beneficial). Practical, current experience with GPU drivers, firmware, and runtime stacks on AI training and inference clusters. Confident with nvidia-smi, DCGM, and XID/error interpretation; able to isolate faults across GPU, baseboard, NIC, and PCIe layers and drive them through diagnosis to RMA. High-performance east-west fabrics. Hands-on experience with RDMA fabrics - InfiniBand and/or RoCE - including link-layer diagnostics (mlxlink, ibdiagnet, or equivalent), transceiver and cabling fault isolation, and understanding of rail-optimised topologies, NVLink/NVSwitch, and NCCL-based performance troubleshooting on multi-node clusters. HPC scheduling. Slurm operations for large multi-GPU jobs - containers via Pyxis/Enroot, MPI, and diagnosing queue, topology, and job failures. Linux systems engineering at scale. Strong command of modern Linux distributions, kernel modules, systemd, networking stack, and filesystem tooling. Proven troubleshooting across compute, storage, and network layers in production. Server hardware and control planes. Comfortable with BMC/Redfish, firmware management, and bare-metal provisioning workflows (MAAS or similar) across large node fleets. Networking fundamentals. Solid grasp of L2/L3, routing, BGP, VLANs, VXLAN, firewalls, and load balancing, with a clear understanding of how east-west cluster traffic differs from north-south. Observability and incident response. Build and use alerting stacks and dashboards (Prometheus/Grafana or similar), interpret metrics and alerts, drive runbooks to resolution, and contribute to SLOs and post-incident reviews. Change and risk judgment. Experience authoring and executing changes in business-critical environments, including risk assessments, customer-impact analysis, and backout plans. SRE-style operations. Write and maintain runbooks, automate diagnostics, and reduce human intervention through scripts and small tools. Automation and Git. Scripting skills in Bash, Python, or equivalent for operational tooling and integrations; experience with infrastructure automation tools (Ansible, Terraform, or similar). Data centre fundamentals. Understanding of how data centres operate - servers, networks, storage, power, and cooling - ideally gained through an operational support background. Leadership. Disciplined, organised, and self-motivated, with the ability to mentor and motivate other engineers, take decisive action, and drive the team and wider organisation to improve. Adaptability. Able to adapt to customer-driven demands, including specialist support outside core hours and travel for onsite work. Nice to Have High-performance storage. Hands-on experience with VAST or comparable AI-optimised storage platforms, or Ceph/parallel filesystems and NFS at scale (multipath, remoteports, nconnect), including diagnosing storage-network interaction and data-path performance issues. OpenStack and fleet operations tooling. OpenStack operations experience (Neutron, Cinder, error triage), plus familiarity with fleet-scale tooling for provisioning, health, and remediation across large GPU estates (MAAS, NetBox, Redfish-driven automation, or similar). Kubernetes. Operating and troubleshooting clusters, including GPU operator stacks and understanding how physical resources are abstracted up the stack. Helpful context for our platform, though not the core of this role. Automation at scale. Automated network configuration with safe, repeatable changes in business-critical environments; GitOps and CI/CD pipelines (GitHub Actions or similar); access and security tooling such as Teleport or Vault in production. Certifications. Relevant GPU/HPC, datacenter architecture, Linux, networking, Kubernetes, cloud, or security certifications (e.g. RHCSA/RHCE, CKA, NVIDIA-certified) are a plus. What We Can Offer You At Nscale, you'll find a collaborative, supportive . click apply for full job details
Raytheon
Sr Principal Real-time Software Engineer
Raytheon Tucson, Arizona
Date Posted: 2026-08-17 Country: United States of America Location: US-AZ-TUCSON- E Hermans Rd BLDG 805 Position Role Type: Onsite U.S. Citizen, U.S. Person, or Immigration Status Requirements: Active and transferable U.S. government issued security clearance is required prior to start date. U.S. citizenship is required, as only U.S. citizens are eligible for a security clearance Security Clearance Type: Secret - Current Security Clearance Status: Active and existing security clearance required on day 1 At RTX, the world largest aerospace and defense company, 185,000 great minds are united by purpose and inspired to make a difference solving the world's most complex problems. With our three market leading businesses, world-class operations and investments in research and development, we offer capabilities and opportunity no one else can. Together, we push the boundaries of known science and find new ways to connect and protect our world. Raytheon brings the strength of more than 100 years of experience and renowned engineering expertise to meet the needs of today's mission and stay ahead of tomorrow's threat. We deliver solutions that help our nation and allies defend freedoms and deter aggression, creating a safer, more secure world. Join us and help shape the future of aerospace and defense. The Software organization develops software applications, including integration and test on missiles, launchers, radars, naval systems, fire control and other complex systems. Our precision software and firmware integrate operating systems, device drivers, networking, and control software to bring together sensor, guidance, and flight control processing features to complete the mission. The Software org is made up of several Centers located across the country, responsible for all aspects of the software development lifecycle. Our 4000+ software engineers design, develop, and build innovative solutions for our customers. Join our fast-paced agile teams on the leading edge of technology. As part of the Software Engineering Directorate's (SWE) Effectors Center (EC) team, you will be an integral part of helping Raytheon further our vision to be the global leader in core and next-generation weapon and security solutions. By any measure, Raytheon is an exciting and rewarding place to work. We pride ourselves on developing mission-driven, world-class talent. The result is a workforce that takes pride in the company and consistently delivers superior solutions. Our Senior Principal Embedded Real-Time Software Developer/Integrator is a technical position that works in an Integrated Product Team (IPT) environment to architect, design, implement, test, debug, and deploy Software for Firmware (FPGA) and Hardware solutions that meet current and next generation autonomous avionics systems' needs. Working with a cross-discipline team, the candidate must have experience developing, testing, and integrating software for edge or embedded devices and/or subsystems (like telecom, medical, IoT, automotive, or robotics) where hardware operation, time critical function, functional reliability, mission assurance, and safety might be major concerns. The successful candidate will work with Product Owners, Chief Engineers, Management and other IPT members using Lean and/or Agile practices to ensure that embedded software is designed and developed to reliably operate toward the intended functions. This position is within the Effectors Center of the Software organization, and is an onsite role located in Tucson, AZ. What You Will Do Architecting, designing, implementing, testing, and debugging integrated embedded real-time software within heterogenous systems composed of firmware and hardware Working within a cross-discipline team to define, refine, and improve product concept, implementation, testability, and guaranteed, measurable quality Teaching, coaching, and mentoring less experienced staff Contributing to proposals as well as preliminary and critical design reviews Ability to obtain program access What You Will Learn Working across a product line in collaboration with other teams Qualifications You Must Have Typically requires a degree in Science, Technology, Engineering or Mathematics (STEM) and a minimum of 10 years of prior relevant experience. Experience including at least two of the following: Embedded C++ Software, Embedded Software Security, Software Architecture Design and Implementation. Experience using embedded Real Time Operating Systems (RTOS) (e.g., Green Hills, Integrity, Wind River VxWorks, Linux, etc.). Experience developing complex systems involving the integration of hardware, firmware, and software. Active and transferable Secret U.S. government issued security clearance is required prior to start date with the ability to obtain program access after start. Qualifications We Prefer Familiarity with rate monotonic theory, practice, and limitations Familiarity with layered architectural principles, and their limitations Familiarity with reading electrical schematics and relating it to software function Familiarity with reading firmware source like VHDL or Verilog Familiarity with assembly language in at least one processor/controller family Experience using lab instruments like power-supplies, digital multi-meters, oscilloscopes, and logic analyzers Experience with developing device drivers for bare-metal and/or OS applications Experience leading engineering teams in delivering systems (of various size) involving the integration of hardware and software. What We Offer Our values drive our actions, behaviors, and performance with a vision for a safer, more connected world. At RTX we value: Safety, Trust, Respect, Accountability, Collaboration, and Innovation. Relocation Offered Based On Eligibility Learn More & Apply Now! Please consider the following role type definition as you apply for this role: Onsite: Employees who are working in Onsite roles will work primarily onsite. This includes all production and maintenance employees, as they are essential to the development of our products. Clearance Information: This position requires a security clearance. DCSA Consolidated Adjudication Services (DCSA CAS), an agency of the Department of Defense, handles and adjudicates the security clearance process. More information about Security Clearances can be found on the US Department of State government website here: Location Information: This position is onsite at our campus in beautiful Tucson, AZ. Tucson has a friendly, caring, and laid-back atmosphere, combined with the innovation and energy of a metropolitan region, and recognized as one of America's 10 Best Small Cities. Surrounded by beautiful mountains, colorful Sonoran Desert landscape and majestic saguaro cacti, Tucson is blessed with some of nature's best work. Tucson is known for its bright blue skies, and with more than 310 sunny days per year, Tucson's fantastic weather lets residents enjoy the outdoors year-round. Virtual Fly Over City of Tucson & Community, YouTube Video Links "Raytheon In Tucson": ,-az-location "Tucson is Awesome": "Winter in Tucson": As part of our commitment to maintaining a secure hiring process, candidates may be asked to attend select steps of the interview process in-person at one of our office locations, regardless of whether the role is designated as on-site, hybrid or remote. The salary range for this role is 132,400 USD - 251,600 USD. The salary range provided is a good faith estimate representative of all experience levels. RTX considers several factors when extending an offer, including but not limited to, the role, function and associated responsibilities, a candidate's work experience, location, education/training, and key skills. Hired applicants may be eligible for benefits, including but not limited to, medical, dental, vision, life insurance, short-term disability, long-term disability, 401(k) match, flexible spending accounts, flexible work schedules, employee assistance program, Employee Scholar Program, parental leave, paid time off, and holidays. Specific benefits are dependent upon the specific business unit as well as whether or not the position is covered by a collective-bargaining agreement. Hired applicants may be eligible for annual short-term and/or long-term incentive compensation programs depending on the level of the position and whether or not it is covered by a collective-bargaining agreement. Payments under these annual programs are not guaranteed and are dependent upon a variety of factors including, but not limited to, individual performance, business unit performance, and/or the company's performance. This role is a U.S.-based role. If the successful candidate resides in a U.S. territory, the appropriate pay structure and benefits will apply. RTX anticipates the application window closing approximately 40 days from the date the notice was posted. However, factors such as candidate flow and business necessity may require RTX to shorten or extend the application window. . click apply for full job details
09/23/2026
Full time
Date Posted: 2026-08-17 Country: United States of America Location: US-AZ-TUCSON- E Hermans Rd BLDG 805 Position Role Type: Onsite U.S. Citizen, U.S. Person, or Immigration Status Requirements: Active and transferable U.S. government issued security clearance is required prior to start date. U.S. citizenship is required, as only U.S. citizens are eligible for a security clearance Security Clearance Type: Secret - Current Security Clearance Status: Active and existing security clearance required on day 1 At RTX, the world largest aerospace and defense company, 185,000 great minds are united by purpose and inspired to make a difference solving the world's most complex problems. With our three market leading businesses, world-class operations and investments in research and development, we offer capabilities and opportunity no one else can. Together, we push the boundaries of known science and find new ways to connect and protect our world. Raytheon brings the strength of more than 100 years of experience and renowned engineering expertise to meet the needs of today's mission and stay ahead of tomorrow's threat. We deliver solutions that help our nation and allies defend freedoms and deter aggression, creating a safer, more secure world. Join us and help shape the future of aerospace and defense. The Software organization develops software applications, including integration and test on missiles, launchers, radars, naval systems, fire control and other complex systems. Our precision software and firmware integrate operating systems, device drivers, networking, and control software to bring together sensor, guidance, and flight control processing features to complete the mission. The Software org is made up of several Centers located across the country, responsible for all aspects of the software development lifecycle. Our 4000+ software engineers design, develop, and build innovative solutions for our customers. Join our fast-paced agile teams on the leading edge of technology. As part of the Software Engineering Directorate's (SWE) Effectors Center (EC) team, you will be an integral part of helping Raytheon further our vision to be the global leader in core and next-generation weapon and security solutions. By any measure, Raytheon is an exciting and rewarding place to work. We pride ourselves on developing mission-driven, world-class talent. The result is a workforce that takes pride in the company and consistently delivers superior solutions. Our Senior Principal Embedded Real-Time Software Developer/Integrator is a technical position that works in an Integrated Product Team (IPT) environment to architect, design, implement, test, debug, and deploy Software for Firmware (FPGA) and Hardware solutions that meet current and next generation autonomous avionics systems' needs. Working with a cross-discipline team, the candidate must have experience developing, testing, and integrating software for edge or embedded devices and/or subsystems (like telecom, medical, IoT, automotive, or robotics) where hardware operation, time critical function, functional reliability, mission assurance, and safety might be major concerns. The successful candidate will work with Product Owners, Chief Engineers, Management and other IPT members using Lean and/or Agile practices to ensure that embedded software is designed and developed to reliably operate toward the intended functions. This position is within the Effectors Center of the Software organization, and is an onsite role located in Tucson, AZ. What You Will Do Architecting, designing, implementing, testing, and debugging integrated embedded real-time software within heterogenous systems composed of firmware and hardware Working within a cross-discipline team to define, refine, and improve product concept, implementation, testability, and guaranteed, measurable quality Teaching, coaching, and mentoring less experienced staff Contributing to proposals as well as preliminary and critical design reviews Ability to obtain program access What You Will Learn Working across a product line in collaboration with other teams Qualifications You Must Have Typically requires a degree in Science, Technology, Engineering or Mathematics (STEM) and a minimum of 10 years of prior relevant experience. Experience including at least two of the following: Embedded C++ Software, Embedded Software Security, Software Architecture Design and Implementation. Experience using embedded Real Time Operating Systems (RTOS) (e.g., Green Hills, Integrity, Wind River VxWorks, Linux, etc.). Experience developing complex systems involving the integration of hardware, firmware, and software. Active and transferable Secret U.S. government issued security clearance is required prior to start date with the ability to obtain program access after start. Qualifications We Prefer Familiarity with rate monotonic theory, practice, and limitations Familiarity with layered architectural principles, and their limitations Familiarity with reading electrical schematics and relating it to software function Familiarity with reading firmware source like VHDL or Verilog Familiarity with assembly language in at least one processor/controller family Experience using lab instruments like power-supplies, digital multi-meters, oscilloscopes, and logic analyzers Experience with developing device drivers for bare-metal and/or OS applications Experience leading engineering teams in delivering systems (of various size) involving the integration of hardware and software. What We Offer Our values drive our actions, behaviors, and performance with a vision for a safer, more connected world. At RTX we value: Safety, Trust, Respect, Accountability, Collaboration, and Innovation. Relocation Offered Based On Eligibility Learn More & Apply Now! Please consider the following role type definition as you apply for this role: Onsite: Employees who are working in Onsite roles will work primarily onsite. This includes all production and maintenance employees, as they are essential to the development of our products. Clearance Information: This position requires a security clearance. DCSA Consolidated Adjudication Services (DCSA CAS), an agency of the Department of Defense, handles and adjudicates the security clearance process. More information about Security Clearances can be found on the US Department of State government website here: Location Information: This position is onsite at our campus in beautiful Tucson, AZ. Tucson has a friendly, caring, and laid-back atmosphere, combined with the innovation and energy of a metropolitan region, and recognized as one of America's 10 Best Small Cities. Surrounded by beautiful mountains, colorful Sonoran Desert landscape and majestic saguaro cacti, Tucson is blessed with some of nature's best work. Tucson is known for its bright blue skies, and with more than 310 sunny days per year, Tucson's fantastic weather lets residents enjoy the outdoors year-round. Virtual Fly Over City of Tucson & Community, YouTube Video Links "Raytheon In Tucson": ,-az-location "Tucson is Awesome": "Winter in Tucson": As part of our commitment to maintaining a secure hiring process, candidates may be asked to attend select steps of the interview process in-person at one of our office locations, regardless of whether the role is designated as on-site, hybrid or remote. The salary range for this role is 132,400 USD - 251,600 USD. The salary range provided is a good faith estimate representative of all experience levels. RTX considers several factors when extending an offer, including but not limited to, the role, function and associated responsibilities, a candidate's work experience, location, education/training, and key skills. Hired applicants may be eligible for benefits, including but not limited to, medical, dental, vision, life insurance, short-term disability, long-term disability, 401(k) match, flexible spending accounts, flexible work schedules, employee assistance program, Employee Scholar Program, parental leave, paid time off, and holidays. Specific benefits are dependent upon the specific business unit as well as whether or not the position is covered by a collective-bargaining agreement. Hired applicants may be eligible for annual short-term and/or long-term incentive compensation programs depending on the level of the position and whether or not it is covered by a collective-bargaining agreement. Payments under these annual programs are not guaranteed and are dependent upon a variety of factors including, but not limited to, individual performance, business unit performance, and/or the company's performance. This role is a U.S.-based role. If the successful candidate resides in a U.S. territory, the appropriate pay structure and benefits will apply. RTX anticipates the application window closing approximately 40 days from the date the notice was posted. However, factors such as candidate flow and business necessity may require RTX to shorten or extend the application window. . click apply for full job details

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board