Description: Transform the Way We Work Are you the type of developer who enjoys solving real business problems, not just writing code? We are looking for a practical, product-minded Business Systems Developer to help us leverage technology, automation, data, and artificial intelligence to improve how our business operates and how our customers experience working with us. This is not a traditional software development role. You will work directly with teams across the organization to understand workflows, identify opportunities for improvement, and build useful software tools that make our company faster, smarter, and easier to do business with. This is a high-impact opportunity for someone who wants to help shape the future of a growing organization through thoughtful application of modern technology. What You'll Do Partner with leaders and employees across departments to understand how work gets done and identify opportunities for improvement. Attack specific business process improvement opportunities identified by senior leadership. Analyze existing workflows and recommend technology-driven solutions that increase efficiency, accuracy, and scalability. Build internal applications, dashboards, integrations, and workflow automations that reduce manual effort and improve visibility. Improve the flow of data between systems and help establish more reliable business processes. Develop customer-facing and customer-supporting tools that enhance speed, communication, and overall customer experience. Utilize software and AI solutions to support customer communications, document processing, knowledge management, reporting, task routing, and decision support. Evaluate emerging technologies and AI tools with a focus on practical application, measurable business value, reliability, and responsible use. Collaborate with technical and non-technical colleagues to ensure solutions meet business needs and drive adoption. Continuously seek opportunities to modernize and improve the organization's technology capabilities. What We're Looking For Experience building software solutions used by internal teams and/or customers. Strong ability to understand business processes and translate operational challenges into practical software solutions. Experience developing web applications, integrations, APIs, databases, and cloud-based solutions. Experience with automation platforms and modern development tools. Familiarity with artificial intelligence technologies and their practical business applications. Excellent communication skills with the ability to work effectively across all levels of the organization. A problem-solving mindset and a bias toward delivering useful, maintainable solutions. Ability to balance technical excellence with business priorities and real-world constraints. Preferred Qualifications Bachelor's degree in Computer Science, Software Engineering, Information Systems, or related field, or equivalent practical experience. Experience developing and supporting cloud-based applications, platforms, and services. Experience with APIs, SQL databases, system integrations, and business applications. Experience implementing automation, workflow, or AI-powered solutions. Experience working in manufacturing, operations, engineering, or project-based environments is a plus. Why This Role Matters This position will play a key role in helping us modernize how we operate. The solutions you develop will directly impact employee productivity, operational efficiency, data quality, decision-making, and customer experience. If you're excited about solving meaningful business challenges with technology and want to see the direct impact of your work, we'd love to hear from you. Get a glimpse of life at Takeform: Find your future with us. Affirmative Action/Equal Opportunity Employer Requirements: Compensation details: 00 Yearly Salary PId9facbf28c96-0343
08/05/2026
Full time
Description: Transform the Way We Work Are you the type of developer who enjoys solving real business problems, not just writing code? We are looking for a practical, product-minded Business Systems Developer to help us leverage technology, automation, data, and artificial intelligence to improve how our business operates and how our customers experience working with us. This is not a traditional software development role. You will work directly with teams across the organization to understand workflows, identify opportunities for improvement, and build useful software tools that make our company faster, smarter, and easier to do business with. This is a high-impact opportunity for someone who wants to help shape the future of a growing organization through thoughtful application of modern technology. What You'll Do Partner with leaders and employees across departments to understand how work gets done and identify opportunities for improvement. Attack specific business process improvement opportunities identified by senior leadership. Analyze existing workflows and recommend technology-driven solutions that increase efficiency, accuracy, and scalability. Build internal applications, dashboards, integrations, and workflow automations that reduce manual effort and improve visibility. Improve the flow of data between systems and help establish more reliable business processes. Develop customer-facing and customer-supporting tools that enhance speed, communication, and overall customer experience. Utilize software and AI solutions to support customer communications, document processing, knowledge management, reporting, task routing, and decision support. Evaluate emerging technologies and AI tools with a focus on practical application, measurable business value, reliability, and responsible use. Collaborate with technical and non-technical colleagues to ensure solutions meet business needs and drive adoption. Continuously seek opportunities to modernize and improve the organization's technology capabilities. What We're Looking For Experience building software solutions used by internal teams and/or customers. Strong ability to understand business processes and translate operational challenges into practical software solutions. Experience developing web applications, integrations, APIs, databases, and cloud-based solutions. Experience with automation platforms and modern development tools. Familiarity with artificial intelligence technologies and their practical business applications. Excellent communication skills with the ability to work effectively across all levels of the organization. A problem-solving mindset and a bias toward delivering useful, maintainable solutions. Ability to balance technical excellence with business priorities and real-world constraints. Preferred Qualifications Bachelor's degree in Computer Science, Software Engineering, Information Systems, or related field, or equivalent practical experience. Experience developing and supporting cloud-based applications, platforms, and services. Experience with APIs, SQL databases, system integrations, and business applications. Experience implementing automation, workflow, or AI-powered solutions. Experience working in manufacturing, operations, engineering, or project-based environments is a plus. Why This Role Matters This position will play a key role in helping us modernize how we operate. The solutions you develop will directly impact employee productivity, operational efficiency, data quality, decision-making, and customer experience. If you're excited about solving meaningful business challenges with technology and want to see the direct impact of your work, we'd love to hear from you. Get a glimpse of life at Takeform: Find your future with us. Affirmative Action/Equal Opportunity Employer Requirements: Compensation details: 00 Yearly Salary PId9facbf28c96-0343
Description: Transform the Way We Work Are you the type of developer who enjoys solving real business problems, not just writing code? We are looking for a practical, product-minded Business Systems Developer to help us leverage technology, automation, data, and artificial intelligence to improve how our business operates and how our customers experience working with us. This is not a traditional software development role. You will work directly with teams across the organization to understand workflows, identify opportunities for improvement, and build useful software tools that make our company faster, smarter, and easier to do business with. This is a high-impact opportunity for someone who wants to help shape the future of a growing organization through thoughtful application of modern technology. What You'll Do Partner with leaders and employees across departments to understand how work gets done and identify opportunities for improvement. Attack specific business process improvement opportunities identified by senior leadership. Analyze existing workflows and recommend technology-driven solutions that increase efficiency, accuracy, and scalability. Build internal applications, dashboards, integrations, and workflow automations that reduce manual effort and improve visibility. Improve the flow of data between systems and help establish more reliable business processes. Develop customer-facing and customer-supporting tools that enhance speed, communication, and overall customer experience. Utilize software and AI solutions to support customer communications, document processing, knowledge management, reporting, task routing, and decision support. Evaluate emerging technologies and AI tools with a focus on practical application, measurable business value, reliability, and responsible use. Collaborate with technical and non-technical colleagues to ensure solutions meet business needs and drive adoption. Continuously seek opportunities to modernize and improve the organization's technology capabilities. What We're Looking For Experience building software solutions used by internal teams and/or customers. Strong ability to understand business processes and translate operational challenges into practical software solutions. Experience developing web applications, integrations, APIs, databases, and cloud-based solutions. Experience with automation platforms and modern development tools. Familiarity with artificial intelligence technologies and their practical business applications. Excellent communication skills with the ability to work effectively across all levels of the organization. A problem-solving mindset and a bias toward delivering useful, maintainable solutions. Ability to balance technical excellence with business priorities and real-world constraints. Preferred Qualifications Bachelor's degree in Computer Science, Software Engineering, Information Systems, or related field, or equivalent practical experience. Experience developing and supporting cloud-based applications, platforms, and services. Experience with APIs, SQL databases, system integrations, and business applications. Experience implementing automation, workflow, or AI-powered solutions. Experience working in manufacturing, operations, engineering, or project-based environments is a plus. Why This Role Matters This position will play a key role in helping us modernize how we operate. The solutions you develop will directly impact employee productivity, operational efficiency, data quality, decision-making, and customer experience. If you're excited about solving meaningful business challenges with technology and want to see the direct impact of your work, we'd love to hear from you. Get a glimpse of life at Takeform: Find your future with us. Affirmative Action/Equal Opportunity Employer Requirements: Compensation details: 00 Yearly Salary PIac965ea5edc3-0343
08/05/2026
Full time
Description: Transform the Way We Work Are you the type of developer who enjoys solving real business problems, not just writing code? We are looking for a practical, product-minded Business Systems Developer to help us leverage technology, automation, data, and artificial intelligence to improve how our business operates and how our customers experience working with us. This is not a traditional software development role. You will work directly with teams across the organization to understand workflows, identify opportunities for improvement, and build useful software tools that make our company faster, smarter, and easier to do business with. This is a high-impact opportunity for someone who wants to help shape the future of a growing organization through thoughtful application of modern technology. What You'll Do Partner with leaders and employees across departments to understand how work gets done and identify opportunities for improvement. Attack specific business process improvement opportunities identified by senior leadership. Analyze existing workflows and recommend technology-driven solutions that increase efficiency, accuracy, and scalability. Build internal applications, dashboards, integrations, and workflow automations that reduce manual effort and improve visibility. Improve the flow of data between systems and help establish more reliable business processes. Develop customer-facing and customer-supporting tools that enhance speed, communication, and overall customer experience. Utilize software and AI solutions to support customer communications, document processing, knowledge management, reporting, task routing, and decision support. Evaluate emerging technologies and AI tools with a focus on practical application, measurable business value, reliability, and responsible use. Collaborate with technical and non-technical colleagues to ensure solutions meet business needs and drive adoption. Continuously seek opportunities to modernize and improve the organization's technology capabilities. What We're Looking For Experience building software solutions used by internal teams and/or customers. Strong ability to understand business processes and translate operational challenges into practical software solutions. Experience developing web applications, integrations, APIs, databases, and cloud-based solutions. Experience with automation platforms and modern development tools. Familiarity with artificial intelligence technologies and their practical business applications. Excellent communication skills with the ability to work effectively across all levels of the organization. A problem-solving mindset and a bias toward delivering useful, maintainable solutions. Ability to balance technical excellence with business priorities and real-world constraints. Preferred Qualifications Bachelor's degree in Computer Science, Software Engineering, Information Systems, or related field, or equivalent practical experience. Experience developing and supporting cloud-based applications, platforms, and services. Experience with APIs, SQL databases, system integrations, and business applications. Experience implementing automation, workflow, or AI-powered solutions. Experience working in manufacturing, operations, engineering, or project-based environments is a plus. Why This Role Matters This position will play a key role in helping us modernize how we operate. The solutions you develop will directly impact employee productivity, operational efficiency, data quality, decision-making, and customer experience. If you're excited about solving meaningful business challenges with technology and want to see the direct impact of your work, we'd love to hear from you. Get a glimpse of life at Takeform: Find your future with us. Affirmative Action/Equal Opportunity Employer Requirements: Compensation details: 00 Yearly Salary PIac965ea5edc3-0343
Join our dynamic AWS team and become a critical guardian of global cloud infrastructure! You'll play a pivotal role in maintaining the heartbeat of the world's most innovative technology platform, ensuring seamless data center operations that power millions of businesses and services worldwide. Data Center Technicians should be willing to work both independently and with a team. Work prioritization, organizational skills, effective communication, and the ability to react quickly are critical to success. In addition to hardware and network repair, candidates will install equipment, create documentation, innovate solutions, and fix problems within the data center space. This team works in an environment that operates 24/7. Traveling within and outside of the regional work area is required. AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we're the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain - and we're looking for talented people who want to help. You'll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You'll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you'll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion. NOTE: PDX is an AWS GovCloud region. As required by our contracts with the federal government, effective February 24, 2026, logical access to the AWS GovCloud region will be restricted to Amazon employees who are U.S. Citizens. (GovCloud may NOT be accessed from outside of the United States). NOTE: Lump sum stipend will be provided to eligible candidates who relocate for this position. is position. Key job responsibilities • Performing server rack installations, hardware break-fix on various components, troubleshooting network issues, and responding to operational incidents that impact service availability, escalating to senior technicians and management as needed/required • Serving as a primary contact point for both internal and external stakeholders, including engineers, software developers, vendors, and contractors • Performing on-call duties and participating in scheduled maintenance and change management activities • Contributing to documentation and process improvement initiatives based on your analysis of operational issues • Helping to train and onboard new team members Physical requirements: • Regularly lift and/or move up to 40 pounds; and participate in group lifts for 41+ pounds • Working in cramped and/or elevated locations • Bending, lifting, stretching, and reaching • Standing and walking for up to 8+ hours a day • Ascending and descending ladders, stairs, and gangways safely and without limitation • Work in a noisy environment • Work shifts longer than 8 hours to support 24/7 operations (covering both night and day shifts). This role supports data centers in various operational stages. The position may be assigned to a facility still under development, which could require travel to established locations for training until your designated site becomes fully operational. Essential Requirements: • Ability to travel to or commute between data center locations as needed • Willingness to temporarily work at alternative sites during training periods or until assigned facility is operational • Travel frequency will vary based on business needs and operational status of assigned facility Candidates must be able to travel (over 50 miles) or commute (less than 50 miles) to active data center sites as required by business needs. A day in the life As a Data Center Technician professional, you have industry-leading technical abilities and demonstrate a breadth of knowledge while you: • Take ownership of technical issues brought by their customer base, engaging other teams when needed to drive resolution • Solve problems at their root and step back to understand the broader context • Maintain service level agreements through the implementation of proactive issue detection and reporting You will be required to work shift work that will include days/nights/weekends/holidays. Traveling within and outside of the regional work area is required. You will be responsible for having a reliable personal vehicle and valid driver's license to travel within the regional work area, as company transportation will not be provided. BASIC QUALIFICATIONS - 1+ years of computer/server hardware troubleshooting or related IT experience - High school or equivalent diploma - Work a flexible schedule including weekends, nights, and holidays PREFERRED QUALIFICATIONS - Knowledge of network design, protocols and Layer 1/2 troubleshooting - Associate's degree in a relevant field (e.g., Computer Science, Networking Engineering), or experience in a relevant field (e.g., Computer Science, Networking Engineering) - Experience managing work and priorities through a ticketing system - Experience in a data center or other critical environment Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The starting pay for this position is listed below. Final starting pay will be based on factors including experience, qualifications, and location. Starting Day 1 of employment, Amazon offers EAP, Mental Health Support, Medical Advice Line, 401(k) matching. Learn more about our benefits at . USA, OR, Umatilla - 27.00 - 48.00 USD hourly
08/05/2026
Full time
Join our dynamic AWS team and become a critical guardian of global cloud infrastructure! You'll play a pivotal role in maintaining the heartbeat of the world's most innovative technology platform, ensuring seamless data center operations that power millions of businesses and services worldwide. Data Center Technicians should be willing to work both independently and with a team. Work prioritization, organizational skills, effective communication, and the ability to react quickly are critical to success. In addition to hardware and network repair, candidates will install equipment, create documentation, innovate solutions, and fix problems within the data center space. This team works in an environment that operates 24/7. Traveling within and outside of the regional work area is required. AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we're the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain - and we're looking for talented people who want to help. You'll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You'll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you'll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion. NOTE: PDX is an AWS GovCloud region. As required by our contracts with the federal government, effective February 24, 2026, logical access to the AWS GovCloud region will be restricted to Amazon employees who are U.S. Citizens. (GovCloud may NOT be accessed from outside of the United States). NOTE: Lump sum stipend will be provided to eligible candidates who relocate for this position. is position. Key job responsibilities • Performing server rack installations, hardware break-fix on various components, troubleshooting network issues, and responding to operational incidents that impact service availability, escalating to senior technicians and management as needed/required • Serving as a primary contact point for both internal and external stakeholders, including engineers, software developers, vendors, and contractors • Performing on-call duties and participating in scheduled maintenance and change management activities • Contributing to documentation and process improvement initiatives based on your analysis of operational issues • Helping to train and onboard new team members Physical requirements: • Regularly lift and/or move up to 40 pounds; and participate in group lifts for 41+ pounds • Working in cramped and/or elevated locations • Bending, lifting, stretching, and reaching • Standing and walking for up to 8+ hours a day • Ascending and descending ladders, stairs, and gangways safely and without limitation • Work in a noisy environment • Work shifts longer than 8 hours to support 24/7 operations (covering both night and day shifts). This role supports data centers in various operational stages. The position may be assigned to a facility still under development, which could require travel to established locations for training until your designated site becomes fully operational. Essential Requirements: • Ability to travel to or commute between data center locations as needed • Willingness to temporarily work at alternative sites during training periods or until assigned facility is operational • Travel frequency will vary based on business needs and operational status of assigned facility Candidates must be able to travel (over 50 miles) or commute (less than 50 miles) to active data center sites as required by business needs. A day in the life As a Data Center Technician professional, you have industry-leading technical abilities and demonstrate a breadth of knowledge while you: • Take ownership of technical issues brought by their customer base, engaging other teams when needed to drive resolution • Solve problems at their root and step back to understand the broader context • Maintain service level agreements through the implementation of proactive issue detection and reporting You will be required to work shift work that will include days/nights/weekends/holidays. Traveling within and outside of the regional work area is required. You will be responsible for having a reliable personal vehicle and valid driver's license to travel within the regional work area, as company transportation will not be provided. BASIC QUALIFICATIONS - 1+ years of computer/server hardware troubleshooting or related IT experience - High school or equivalent diploma - Work a flexible schedule including weekends, nights, and holidays PREFERRED QUALIFICATIONS - Knowledge of network design, protocols and Layer 1/2 troubleshooting - Associate's degree in a relevant field (e.g., Computer Science, Networking Engineering), or experience in a relevant field (e.g., Computer Science, Networking Engineering) - Experience managing work and priorities through a ticketing system - Experience in a data center or other critical environment Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner. The starting pay for this position is listed below. Final starting pay will be based on factors including experience, qualifications, and location. Starting Day 1 of employment, Amazon offers EAP, Mental Health Support, Medical Advice Line, 401(k) matching. Learn more about our benefits at . USA, OR, Umatilla - 27.00 - 48.00 USD hourly
Job Description Job Description We are looking for a senior ML infrastructure engineer to build and evolve the systems that support model training, deployment, and production usage. This role sits at the intersection of software engineering, infrastructure, ML workflows, and developer experience. The work focuses on production-quality ML systems: reliability, scalability, observability, and usability for the team. You will work on training infrastructure, deployment workflows, model serving, platform tooling, automation, and production reliability. Training deployment is a core focus, and experience with inference deployment is a strong plus. We care about engineering judgment, technical depth, communication, and the ability to turn messy ML workflows into stable platform capabilities. What You Will Own Design, build, and evolve infrastructure for ML training workflows, training deployment, experiment execution, and production handoff. Build and maintain deployment paths for models, jobs, services, and supporting infrastructure across development and production environments. Improve reliability, scalability, observability, and developer experience for ML workflows and platform tools. Define interfaces, automation, metadata, artifacts, configuration, environment management, and lifecycle boundaries for ML systems. Collaborate with research, product, data, and engineering partners to translate incomplete ML workflow needs into maintainable systems. Support production usage by building clear operational tooling, debugging paths, and safe rollout mechanisms. What We Look For Strong software engineering and infrastructure fundamentals, with experience owning production or near-production systems. Practical experience with PyTorch and ML training workflows, including job orchestration, compute environments, artifact management, and deployment automation. Solid understanding of heterogeneous computing and high-performance computing, especially for ML training or serving workloads. Good understanding of model lifecycle concerns: data, configs, checkpoints, artifacts, reproducibility, rollout, rollback, and observability. Ability to build reliable platform abstractions without hiding the important details ML practitioners need to control. Clear technical and product sense: you can prioritize platform work that unlocks real training or deployment velocity. High standards for engineering quality, including tests, documentation, debugging tools, and maintainable system design. Tech Stack You May Work With Python PyTorch Heterogeneous computing and high-performance computing ML training pipelines, job orchestration, compute scheduling, containers, and deployment automation Model artifacts, metadata, storage, experiment tracking, and configuration systems Model serving, inference deployment, APIs, queues, and observability tools Docker, CI/CD, cloud infrastructure, GPUs, and internal platform tooling Bonus Points Experience building training deployment systems, model release workflows, or ML platform tooling for research and production teams. Experience with inference deployment, model serving, online/offline evaluation, performance tuning, or rollout safety. Experience with distributed training, GPU infrastructure, workload scheduling, artifact/version management, or reproducibility tooling. Understanding of CUDA, GPU architecture, or low-level performance optimization. Experience with open-source inference and serving frameworks such as vLLM, TensorRT, Triton, or similar systems. Experience migrating ad hoc notebooks, scripts, or manual ML processes into reliable platform workflows. This Role May Not Be a Fit If You mainly want to train models personally and do not enjoy building infrastructure for others to use. You are comfortable with manual ML workflows and do not care about reproducibility, deployment, or operational quality. You prefer narrow implementation tasks and do not want to reason about system boundaries, platform UX, or long-term maintenance. You over-abstract ML workflows without understanding where researchers and engineers need control and visibility. Why This Role Matters ML teams move faster when training, deployment, and production usage are supported by reliable infrastructure instead of scattered scripts and manual processes. This role will directly shape how models move from experimentation to production, how safely they are deployed, and how efficiently the team can iterate. For the right person, it is a high-ownership platform role with deep impact on both engineering quality and ML velocity. What We Would Like to See When You Apply ML infrastructure, training platforms, deployment systems, or model serving systems you have owned. Examples of how you improved training reliability, deployment velocity, reproducibility, observability, or operational safety. Cases where you turned messy ML workflows into maintainable tools, services, or platform abstractions. Examples that show your technical judgment, communication, and ability to work across research and engineering needs.
08/05/2026
Full time
Job Description Job Description We are looking for a senior ML infrastructure engineer to build and evolve the systems that support model training, deployment, and production usage. This role sits at the intersection of software engineering, infrastructure, ML workflows, and developer experience. The work focuses on production-quality ML systems: reliability, scalability, observability, and usability for the team. You will work on training infrastructure, deployment workflows, model serving, platform tooling, automation, and production reliability. Training deployment is a core focus, and experience with inference deployment is a strong plus. We care about engineering judgment, technical depth, communication, and the ability to turn messy ML workflows into stable platform capabilities. What You Will Own Design, build, and evolve infrastructure for ML training workflows, training deployment, experiment execution, and production handoff. Build and maintain deployment paths for models, jobs, services, and supporting infrastructure across development and production environments. Improve reliability, scalability, observability, and developer experience for ML workflows and platform tools. Define interfaces, automation, metadata, artifacts, configuration, environment management, and lifecycle boundaries for ML systems. Collaborate with research, product, data, and engineering partners to translate incomplete ML workflow needs into maintainable systems. Support production usage by building clear operational tooling, debugging paths, and safe rollout mechanisms. What We Look For Strong software engineering and infrastructure fundamentals, with experience owning production or near-production systems. Practical experience with PyTorch and ML training workflows, including job orchestration, compute environments, artifact management, and deployment automation. Solid understanding of heterogeneous computing and high-performance computing, especially for ML training or serving workloads. Good understanding of model lifecycle concerns: data, configs, checkpoints, artifacts, reproducibility, rollout, rollback, and observability. Ability to build reliable platform abstractions without hiding the important details ML practitioners need to control. Clear technical and product sense: you can prioritize platform work that unlocks real training or deployment velocity. High standards for engineering quality, including tests, documentation, debugging tools, and maintainable system design. Tech Stack You May Work With Python PyTorch Heterogeneous computing and high-performance computing ML training pipelines, job orchestration, compute scheduling, containers, and deployment automation Model artifacts, metadata, storage, experiment tracking, and configuration systems Model serving, inference deployment, APIs, queues, and observability tools Docker, CI/CD, cloud infrastructure, GPUs, and internal platform tooling Bonus Points Experience building training deployment systems, model release workflows, or ML platform tooling for research and production teams. Experience with inference deployment, model serving, online/offline evaluation, performance tuning, or rollout safety. Experience with distributed training, GPU infrastructure, workload scheduling, artifact/version management, or reproducibility tooling. Understanding of CUDA, GPU architecture, or low-level performance optimization. Experience with open-source inference and serving frameworks such as vLLM, TensorRT, Triton, or similar systems. Experience migrating ad hoc notebooks, scripts, or manual ML processes into reliable platform workflows. This Role May Not Be a Fit If You mainly want to train models personally and do not enjoy building infrastructure for others to use. You are comfortable with manual ML workflows and do not care about reproducibility, deployment, or operational quality. You prefer narrow implementation tasks and do not want to reason about system boundaries, platform UX, or long-term maintenance. You over-abstract ML workflows without understanding where researchers and engineers need control and visibility. Why This Role Matters ML teams move faster when training, deployment, and production usage are supported by reliable infrastructure instead of scattered scripts and manual processes. This role will directly shape how models move from experimentation to production, how safely they are deployed, and how efficiently the team can iterate. For the right person, it is a high-ownership platform role with deep impact on both engineering quality and ML velocity. What We Would Like to See When You Apply ML infrastructure, training platforms, deployment systems, or model serving systems you have owned. Examples of how you improved training reliability, deployment velocity, reproducibility, observability, or operational safety. Cases where you turned messy ML workflows into maintainable tools, services, or platform abstractions. Examples that show your technical judgment, communication, and ability to work across research and engineering needs.
Job Description Job Description At Sonatus, we're driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can't keep pace with consumer expectations shaped by the mobile industry-where features evolve rapidly, update seamlessly, and improve continuously. That's why leading OEMs trust Sonatus to accelerate this shift. Our technology is already in production across more than 8 million vehicles on the road today and rapidly expanding. Headquartered in Sunnyvale, CA, with 250+ employees worldwide, Sonatus combines the agility of a fast-growing company with the scale and impact of an established partner. Backed by strong funding and proven by global deployment, we're solving some of the most interesting and complex challenges in the industry. Join us and help redefine what's possible as we shape the future of mobility. Role Summary: Sonatus is a global leader in the automotive industry, providing key technologies that enable intelligent AI-defined vehicles. Our solutions are already on the road with millions of vehicles, and we are quickly expanding our offerings for production-grade AI on the Edge. We are looking for a great Senior Staff AI Engineer to join our seasoned AI team and lead the development of Edge AI for in-vehicle self-aware health monitoring and prediction. In this role, you will build and deploy AI models that analyze continuous data generated in the vehicle during the day-to-day operation, including system logs, traces, and vehicle internal signals (Ethernet and CAN) to detect and predict the health of different sub-systems and anticipate failures in real-time. You will own the end-to-end ML pipeline-from data ingestion and model training to deployment on resource-constrained edge devices and model optimization. You will work in a fast-paced startup environment where your code will directly impact fleet reliability and build the next generation of the self-aware vehicle. You will be expected to collaborate with other leading developers who have a deep understanding and expertise of vehicle software and systems, and other AI developers working on MLOps and integration of AI models on vehicles expected to be on the road today. Expect to experiment with cutting-edge model architectures and best-in-class development tools. This is a hybrid role out of our Sunnyvale, CA, where you will be expected to work in our office 3 days a week. Responsibilities: Build and train AI Edge models (e.g., Transformers, LLM, CNN, LSTM, Trees) to process unstructured application logs, kernel traces, and multi-modalities. Integrating ML flows, including cloud-based LLM APIs (Gemini, OpenAI, Claude), with emphasis on synthetic data creation. Develop algorithms to automatically cluster log patterns and detect software regressions, race conditions, or crash precursors. Design unsupervised and supervised learning models (e.g., Autoencoders, Isolation Forests) to monitor time-series data from CAN bus and on-board sensors. Implement logic to correlate signal anomalies (e.g., ADAS drifts, sensor spikes, latency jitters) across different modalities with system events to identify root causes. Port and optimize PyTorch/TensorFlow models into production-grade models for execution on CPU/GPU-bound targets or embedded NPUs. Apply quantization, pruning, distillation, and memory optimization to ensure models run within strict RAM/Flash budgets. Define the data strategy for on-device filtering: pre-processing on device and decide which data is processed locally versus processed in the cloud. Lead the architecture for the edge ML pipeline and mentor junior engineers on best practices for embedded AI. Requirements: Bachelor's degree in Computer Science, Electrical Engineering, Software Engineering, or a related field. 10+ years in Machine Learning Engineering, with 3+ years focused on Edge AI or Embedded Systems. Proven experience mentoring junior engineers in software development. Expert Python (for training) and decent working knowledge of modern C++ (C+/17 for inference). Deep proficiency with PyTorch or TensorFlow, and experience with inference engines like ONNX, TFLite, or TVM. Experience with NLP techniques for textual data parsing, sequence modeling (RNN/GRU), vector store, or lightweight LLMs/SLMs. Experience with libraries like scikit-learn, tslearn, or statsmodels for anomaly detection on sensor data. Proven ability to lead technical projects from concept to production in an ambiguous, fast-paced environment. Ability to communicate with stakeholders and articulate trade-offs. Experience deploying to Edge environments (e.g., ARM-based), managing memory manually, and working with limited compute resources. Candidates with a strong Computer Vision (CV) / ADAS track record are highly encouraged to apply! Desired Skills: MS/PhD in Computer Science, Engineering, or related fields. Familiarity with Edge systems and preferably automotive formats (CAN, DBC, UDS, SOME/IP, or MQTT. Understanding of Linux/QNX kernel logs (dmesg), process states, and OS-level debugging. Experience with NVIDIA TensorRT, Qualcomm SNPE. Sunnyvale HQ Benefits & Perks Offered: Health care plan (Medical, Dental & Vision) Flexible and Dependent Care Expense program Retirement plan (401k) Life Insurance (Basic, Voluntary & AD&D) Unlimited paid time off per year, 14+ paid holidays Hybrid office work arrangement Complimentary lunches, snacks, and beverages during on-site working days Wellness benefit allowance Phone & Internet reimbursement Computer Accessory Allowance The posted salary range is a general guideline and represents a good faith estimate of what Sonatus ("Company") could reasonably expect to pay for a base salary for this position. The pay offered to a selected candidate will be determined based on factors such as (but not limited to) the scope and responsibilities of the position, the qualifications of the selected candidate, departmental budget availability, geographic location and external market pay for comparable jobs. The Company reserves the right to modify this range in the future, as needed, as market conditions change. Base Salary Pay Range $227,000-$300,000 USD
08/05/2026
Full time
Job Description Job Description At Sonatus, we're driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can't keep pace with consumer expectations shaped by the mobile industry-where features evolve rapidly, update seamlessly, and improve continuously. That's why leading OEMs trust Sonatus to accelerate this shift. Our technology is already in production across more than 8 million vehicles on the road today and rapidly expanding. Headquartered in Sunnyvale, CA, with 250+ employees worldwide, Sonatus combines the agility of a fast-growing company with the scale and impact of an established partner. Backed by strong funding and proven by global deployment, we're solving some of the most interesting and complex challenges in the industry. Join us and help redefine what's possible as we shape the future of mobility. Role Summary: Sonatus is a global leader in the automotive industry, providing key technologies that enable intelligent AI-defined vehicles. Our solutions are already on the road with millions of vehicles, and we are quickly expanding our offerings for production-grade AI on the Edge. We are looking for a great Senior Staff AI Engineer to join our seasoned AI team and lead the development of Edge AI for in-vehicle self-aware health monitoring and prediction. In this role, you will build and deploy AI models that analyze continuous data generated in the vehicle during the day-to-day operation, including system logs, traces, and vehicle internal signals (Ethernet and CAN) to detect and predict the health of different sub-systems and anticipate failures in real-time. You will own the end-to-end ML pipeline-from data ingestion and model training to deployment on resource-constrained edge devices and model optimization. You will work in a fast-paced startup environment where your code will directly impact fleet reliability and build the next generation of the self-aware vehicle. You will be expected to collaborate with other leading developers who have a deep understanding and expertise of vehicle software and systems, and other AI developers working on MLOps and integration of AI models on vehicles expected to be on the road today. Expect to experiment with cutting-edge model architectures and best-in-class development tools. This is a hybrid role out of our Sunnyvale, CA, where you will be expected to work in our office 3 days a week. Responsibilities: Build and train AI Edge models (e.g., Transformers, LLM, CNN, LSTM, Trees) to process unstructured application logs, kernel traces, and multi-modalities. Integrating ML flows, including cloud-based LLM APIs (Gemini, OpenAI, Claude), with emphasis on synthetic data creation. Develop algorithms to automatically cluster log patterns and detect software regressions, race conditions, or crash precursors. Design unsupervised and supervised learning models (e.g., Autoencoders, Isolation Forests) to monitor time-series data from CAN bus and on-board sensors. Implement logic to correlate signal anomalies (e.g., ADAS drifts, sensor spikes, latency jitters) across different modalities with system events to identify root causes. Port and optimize PyTorch/TensorFlow models into production-grade models for execution on CPU/GPU-bound targets or embedded NPUs. Apply quantization, pruning, distillation, and memory optimization to ensure models run within strict RAM/Flash budgets. Define the data strategy for on-device filtering: pre-processing on device and decide which data is processed locally versus processed in the cloud. Lead the architecture for the edge ML pipeline and mentor junior engineers on best practices for embedded AI. Requirements: Bachelor's degree in Computer Science, Electrical Engineering, Software Engineering, or a related field. 10+ years in Machine Learning Engineering, with 3+ years focused on Edge AI or Embedded Systems. Proven experience mentoring junior engineers in software development. Expert Python (for training) and decent working knowledge of modern C++ (C+/17 for inference). Deep proficiency with PyTorch or TensorFlow, and experience with inference engines like ONNX, TFLite, or TVM. Experience with NLP techniques for textual data parsing, sequence modeling (RNN/GRU), vector store, or lightweight LLMs/SLMs. Experience with libraries like scikit-learn, tslearn, or statsmodels for anomaly detection on sensor data. Proven ability to lead technical projects from concept to production in an ambiguous, fast-paced environment. Ability to communicate with stakeholders and articulate trade-offs. Experience deploying to Edge environments (e.g., ARM-based), managing memory manually, and working with limited compute resources. Candidates with a strong Computer Vision (CV) / ADAS track record are highly encouraged to apply! Desired Skills: MS/PhD in Computer Science, Engineering, or related fields. Familiarity with Edge systems and preferably automotive formats (CAN, DBC, UDS, SOME/IP, or MQTT. Understanding of Linux/QNX kernel logs (dmesg), process states, and OS-level debugging. Experience with NVIDIA TensorRT, Qualcomm SNPE. Sunnyvale HQ Benefits & Perks Offered: Health care plan (Medical, Dental & Vision) Flexible and Dependent Care Expense program Retirement plan (401k) Life Insurance (Basic, Voluntary & AD&D) Unlimited paid time off per year, 14+ paid holidays Hybrid office work arrangement Complimentary lunches, snacks, and beverages during on-site working days Wellness benefit allowance Phone & Internet reimbursement Computer Accessory Allowance The posted salary range is a general guideline and represents a good faith estimate of what Sonatus ("Company") could reasonably expect to pay for a base salary for this position. The pay offered to a selected candidate will be determined based on factors such as (but not limited to) the scope and responsibilities of the position, the qualifications of the selected candidate, departmental budget availability, geographic location and external market pay for comparable jobs. The Company reserves the right to modify this range in the future, as needed, as market conditions change. Base Salary Pay Range $227,000-$300,000 USD
Job Description Job Description At Sonatus, we're driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can't keep pace with consumer expectations shaped by the mobile industry-where features evolve rapidly, update seamlessly, and improve continuously. That's why leading OEMs trust Sonatus to accelerate this shift. Our technology is already in production across more than 8 million vehicles on the road today and rapidly expanding. Headquartered in Sunnyvale, CA, with 250+ employees worldwide, Sonatus combines the agility of a fast-growing company with the scale and impact of an established partner. Backed by strong funding and proven by global deployment, we're solving some of the most interesting and complex challenges in the industry. Join us and help redefine what's possible as we shape the future of mobility. Role Summary: Sonatus is a global leader in the automotive industry, providing key technologies that enable intelligent AI-defined vehicles. Our solutions are already on the road with millions of vehicles, and we are quickly expanding our offerings for production-grade AI on the Edge. We are looking for a great Senior Staff AI Engineer to join our seasoned AI team and lead the development of Edge AI for in-vehicle self-aware health monitoring and prediction. In this role, you will build and deploy AI models that analyze continuous data generated in the vehicle during the day-to-day operation, including system logs, traces, and vehicle internal signals (Ethernet and CAN) to detect and predict the health of different sub-systems and anticipate failures in real-time. You will own the end-to-end ML pipeline-from data ingestion and model training to deployment on resource-constrained edge devices and model optimization. You will work in a fast-paced startup environment where your code will directly impact fleet reliability and build the next generation of the self-aware vehicle. You will be expected to collaborate with other leading developers who have a deep understanding and expertise of vehicle software and systems, and other AI developers working on MLOps and integration of AI models on vehicles expected to be on the road today. Expect to experiment with cutting-edge model architectures and best-in-class development tools. This is a hybrid role out of our Sunnyvale, CA, where you will be expected to work in our office 3 days a week. Responsibilities: Build and train AI Edge models (e.g., Transformers, LLM, CNN, LSTM, Trees) to process unstructured application logs, kernel traces, and multi-modalities. Integrating ML flows, including cloud-based LLM APIs (Gemini, OpenAI, Claude), with emphasis on synthetic data creation. Develop algorithms to automatically cluster log patterns and detect software regressions, race conditions, or crash precursors. Design unsupervised and supervised learning models (e.g., Autoencoders, Isolation Forests) to monitor time-series data from CAN bus and on-board sensors. Implement logic to correlate signal anomalies (e.g., ADAS drifts, sensor spikes, latency jitters) across different modalities with system events to identify root causes. Port and optimize PyTorch/TensorFlow models into production-grade models for execution on CPU/GPU-bound targets or embedded NPUs. Apply quantization, pruning, distillation, and memory optimization to ensure models run within strict RAM/Flash budgets. Define the data strategy for on-device filtering: pre-processing on device and decide which data is processed locally versus processed in the cloud. Lead the architecture for the edge ML pipeline and mentor junior engineers on best practices for embedded AI. Requirements: Bachelor's degree in Computer Science, Electrical Engineering, Software Engineering, or a related field. 10+ years in Machine Learning Engineering, with 3+ years focused on Edge AI or Embedded Systems. Proven experience mentoring junior engineers in software development. Expert Python (for training) and decent working knowledge of modern C++ (C+/17 for inference). Deep proficiency with PyTorch or TensorFlow, and experience with inference engines like ONNX, TFLite, or TVM. Experience with NLP techniques for textual data parsing, sequence modeling (RNN/GRU), vector store, or lightweight LLMs/SLMs. Experience with libraries like scikit-learn, tslearn, or statsmodels for anomaly detection on sensor data. Proven ability to lead technical projects from concept to production in an ambiguous, fast-paced environment. Ability to communicate with stakeholders and articulate trade-offs. Experience deploying to Edge environments (e.g., ARM-based), managing memory manually, and working with limited compute resources. Candidates with a strong Computer Vision (CV) / ADAS track record are highly encouraged to apply! Desired Skills: MS/PhD in Computer Science, Engineering, or related fields. Familiarity with Edge systems and preferably automotive formats (CAN, DBC, UDS, SOME/IP, or MQTT. Understanding of Linux/QNX kernel logs (dmesg), process states, and OS-level debugging. Experience with NVIDIA TensorRT, Qualcomm SNPE. Sunnyvale HQ Benefits & Perks Offered: Health care plan (Medical, Dental & Vision) Flexible and Dependent Care Expense program Retirement plan (401k) Life Insurance (Basic, Voluntary & AD&D) Unlimited paid time off per year, 14+ paid holidays Hybrid office work arrangement Complimentary lunches, snacks, and beverages during on-site working days Wellness benefit allowance Phone & Internet reimbursement Computer Accessory Allowance The posted salary range is a general guideline and represents a good faith estimate of what Sonatus ("Company") could reasonably expect to pay for a base salary for this position. The pay offered to a selected candidate will be determined based on factors such as (but not limited to) the scope and responsibilities of the position, the qualifications of the selected candidate, departmental budget availability, geographic location and external market pay for comparable jobs. The Company reserves the right to modify this range in the future, as needed, as market conditions change. Base Salary Pay Range $227,000-$300,000 USD
08/05/2026
Full time
Job Description Job Description At Sonatus, we're driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can't keep pace with consumer expectations shaped by the mobile industry-where features evolve rapidly, update seamlessly, and improve continuously. That's why leading OEMs trust Sonatus to accelerate this shift. Our technology is already in production across more than 8 million vehicles on the road today and rapidly expanding. Headquartered in Sunnyvale, CA, with 250+ employees worldwide, Sonatus combines the agility of a fast-growing company with the scale and impact of an established partner. Backed by strong funding and proven by global deployment, we're solving some of the most interesting and complex challenges in the industry. Join us and help redefine what's possible as we shape the future of mobility. Role Summary: Sonatus is a global leader in the automotive industry, providing key technologies that enable intelligent AI-defined vehicles. Our solutions are already on the road with millions of vehicles, and we are quickly expanding our offerings for production-grade AI on the Edge. We are looking for a great Senior Staff AI Engineer to join our seasoned AI team and lead the development of Edge AI for in-vehicle self-aware health monitoring and prediction. In this role, you will build and deploy AI models that analyze continuous data generated in the vehicle during the day-to-day operation, including system logs, traces, and vehicle internal signals (Ethernet and CAN) to detect and predict the health of different sub-systems and anticipate failures in real-time. You will own the end-to-end ML pipeline-from data ingestion and model training to deployment on resource-constrained edge devices and model optimization. You will work in a fast-paced startup environment where your code will directly impact fleet reliability and build the next generation of the self-aware vehicle. You will be expected to collaborate with other leading developers who have a deep understanding and expertise of vehicle software and systems, and other AI developers working on MLOps and integration of AI models on vehicles expected to be on the road today. Expect to experiment with cutting-edge model architectures and best-in-class development tools. This is a hybrid role out of our Sunnyvale, CA, where you will be expected to work in our office 3 days a week. Responsibilities: Build and train AI Edge models (e.g., Transformers, LLM, CNN, LSTM, Trees) to process unstructured application logs, kernel traces, and multi-modalities. Integrating ML flows, including cloud-based LLM APIs (Gemini, OpenAI, Claude), with emphasis on synthetic data creation. Develop algorithms to automatically cluster log patterns and detect software regressions, race conditions, or crash precursors. Design unsupervised and supervised learning models (e.g., Autoencoders, Isolation Forests) to monitor time-series data from CAN bus and on-board sensors. Implement logic to correlate signal anomalies (e.g., ADAS drifts, sensor spikes, latency jitters) across different modalities with system events to identify root causes. Port and optimize PyTorch/TensorFlow models into production-grade models for execution on CPU/GPU-bound targets or embedded NPUs. Apply quantization, pruning, distillation, and memory optimization to ensure models run within strict RAM/Flash budgets. Define the data strategy for on-device filtering: pre-processing on device and decide which data is processed locally versus processed in the cloud. Lead the architecture for the edge ML pipeline and mentor junior engineers on best practices for embedded AI. Requirements: Bachelor's degree in Computer Science, Electrical Engineering, Software Engineering, or a related field. 10+ years in Machine Learning Engineering, with 3+ years focused on Edge AI or Embedded Systems. Proven experience mentoring junior engineers in software development. Expert Python (for training) and decent working knowledge of modern C++ (C+/17 for inference). Deep proficiency with PyTorch or TensorFlow, and experience with inference engines like ONNX, TFLite, or TVM. Experience with NLP techniques for textual data parsing, sequence modeling (RNN/GRU), vector store, or lightweight LLMs/SLMs. Experience with libraries like scikit-learn, tslearn, or statsmodels for anomaly detection on sensor data. Proven ability to lead technical projects from concept to production in an ambiguous, fast-paced environment. Ability to communicate with stakeholders and articulate trade-offs. Experience deploying to Edge environments (e.g., ARM-based), managing memory manually, and working with limited compute resources. Candidates with a strong Computer Vision (CV) / ADAS track record are highly encouraged to apply! Desired Skills: MS/PhD in Computer Science, Engineering, or related fields. Familiarity with Edge systems and preferably automotive formats (CAN, DBC, UDS, SOME/IP, or MQTT. Understanding of Linux/QNX kernel logs (dmesg), process states, and OS-level debugging. Experience with NVIDIA TensorRT, Qualcomm SNPE. Sunnyvale HQ Benefits & Perks Offered: Health care plan (Medical, Dental & Vision) Flexible and Dependent Care Expense program Retirement plan (401k) Life Insurance (Basic, Voluntary & AD&D) Unlimited paid time off per year, 14+ paid holidays Hybrid office work arrangement Complimentary lunches, snacks, and beverages during on-site working days Wellness benefit allowance Phone & Internet reimbursement Computer Accessory Allowance The posted salary range is a general guideline and represents a good faith estimate of what Sonatus ("Company") could reasonably expect to pay for a base salary for this position. The pay offered to a selected candidate will be determined based on factors such as (but not limited to) the scope and responsibilities of the position, the qualifications of the selected candidate, departmental budget availability, geographic location and external market pay for comparable jobs. The Company reserves the right to modify this range in the future, as needed, as market conditions change. Base Salary Pay Range $227,000-$300,000 USD
Job Description Job Description At Sonatus, we're driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can't keep pace with consumer expectations shaped by the mobile industry-where features evolve rapidly, update seamlessly, and improve continuously. That's why leading OEMs trust Sonatus to accelerate this shift. Our technology is already in production across more than 8 million vehicles on the road today and rapidly expanding. Headquartered in Sunnyvale, CA, with 250+ employees worldwide, Sonatus combines the agility of a fast-growing company with the scale and impact of an established partner. Backed by strong funding and proven by global deployment, we're solving some of the most interesting and complex challenges in the industry. Join us and help redefine what's possible as we shape the future of mobility. Role Summary: Sonatus is a global leader in the automotive industry, providing key technologies that enable intelligent AI-defined vehicles. Our solutions are already on the road with millions of vehicles, and we are quickly expanding our offerings for production-grade AI on the Edge. We are looking for a great Senior Staff AI Engineer to join our seasoned AI team and lead the development of Edge AI for in-vehicle self-aware health monitoring and prediction. In this role, you will build and deploy AI models that analyze continuous data generated in the vehicle during the day-to-day operation, including system logs, traces, and vehicle internal signals (Ethernet and CAN) to detect and predict the health of different sub-systems and anticipate failures in real-time. You will own the end-to-end ML pipeline-from data ingestion and model training to deployment on resource-constrained edge devices and model optimization. You will work in a fast-paced startup environment where your code will directly impact fleet reliability and build the next generation of the self-aware vehicle. You will be expected to collaborate with other leading developers who have a deep understanding and expertise of vehicle software and systems, and other AI developers working on MLOps and integration of AI models on vehicles expected to be on the road today. Expect to experiment with cutting-edge model architectures and best-in-class development tools. This is a hybrid role out of our Sunnyvale, CA, where you will be expected to work in our office 3 days a week. Responsibilities: Build and train AI Edge models (e.g., Transformers, LLM, CNN, LSTM, Trees) to process unstructured application logs, kernel traces, and multi-modalities. Integrating ML flows, including cloud-based LLM APIs (Gemini, OpenAI, Claude), with emphasis on synthetic data creation. Develop algorithms to automatically cluster log patterns and detect software regressions, race conditions, or crash precursors. Design unsupervised and supervised learning models (e.g., Autoencoders, Isolation Forests) to monitor time-series data from CAN bus and on-board sensors. Implement logic to correlate signal anomalies (e.g., ADAS drifts, sensor spikes, latency jitters) across different modalities with system events to identify root causes. Port and optimize PyTorch/TensorFlow models into production-grade models for execution on CPU/GPU-bound targets or embedded NPUs. Apply quantization, pruning, distillation, and memory optimization to ensure models run within strict RAM/Flash budgets. Define the data strategy for on-device filtering: pre-processing on device and decide which data is processed locally versus processed in the cloud. Lead the architecture for the edge ML pipeline and mentor junior engineers on best practices for embedded AI. Requirements: Bachelor's degree in Computer Science, Electrical Engineering, Software Engineering, or a related field. 10+ years in Machine Learning Engineering, with 3+ years focused on Edge AI or Embedded Systems. Proven experience mentoring junior engineers in software development. Expert Python (for training) and decent working knowledge of modern C++ (C+/17 for inference). Deep proficiency with PyTorch or TensorFlow, and experience with inference engines like ONNX, TFLite, or TVM. Experience with NLP techniques for textual data parsing, sequence modeling (RNN/GRU), vector store, or lightweight LLMs/SLMs. Experience with libraries like scikit-learn, tslearn, or statsmodels for anomaly detection on sensor data. Proven ability to lead technical projects from concept to production in an ambiguous, fast-paced environment. Ability to communicate with stakeholders and articulate trade-offs. Experience deploying to Edge environments (e.g., ARM-based), managing memory manually, and working with limited compute resources. Candidates with a strong Computer Vision (CV) / ADAS track record are highly encouraged to apply! Desired Skills: MS/PhD in Computer Science, Engineering, or related fields. Familiarity with Edge systems and preferably automotive formats (CAN, DBC, UDS, SOME/IP, or MQTT. Understanding of Linux/QNX kernel logs (dmesg), process states, and OS-level debugging. Experience with NVIDIA TensorRT, Qualcomm SNPE. Sunnyvale HQ Benefits & Perks Offered: Health care plan (Medical, Dental & Vision) Flexible and Dependent Care Expense program Retirement plan (401k) Life Insurance (Basic, Voluntary & AD&D) Unlimited paid time off per year, 14+ paid holidays Hybrid office work arrangement Complimentary lunches, snacks, and beverages during on-site working days Wellness benefit allowance Phone & Internet reimbursement Computer Accessory Allowance The posted salary range is a general guideline and represents a good faith estimate of what Sonatus ("Company") could reasonably expect to pay for a base salary for this position. The pay offered to a selected candidate will be determined based on factors such as (but not limited to) the scope and responsibilities of the position, the qualifications of the selected candidate, departmental budget availability, geographic location and external market pay for comparable jobs. The Company reserves the right to modify this range in the future, as needed, as market conditions change. Base Salary Pay Range $227,000-$300,000 USD
08/05/2026
Full time
Job Description Job Description At Sonatus, we're driving the transformation to AI-enabled software-defined vehicles. Traditional automotive software methods can't keep pace with consumer expectations shaped by the mobile industry-where features evolve rapidly, update seamlessly, and improve continuously. That's why leading OEMs trust Sonatus to accelerate this shift. Our technology is already in production across more than 8 million vehicles on the road today and rapidly expanding. Headquartered in Sunnyvale, CA, with 250+ employees worldwide, Sonatus combines the agility of a fast-growing company with the scale and impact of an established partner. Backed by strong funding and proven by global deployment, we're solving some of the most interesting and complex challenges in the industry. Join us and help redefine what's possible as we shape the future of mobility. Role Summary: Sonatus is a global leader in the automotive industry, providing key technologies that enable intelligent AI-defined vehicles. Our solutions are already on the road with millions of vehicles, and we are quickly expanding our offerings for production-grade AI on the Edge. We are looking for a great Senior Staff AI Engineer to join our seasoned AI team and lead the development of Edge AI for in-vehicle self-aware health monitoring and prediction. In this role, you will build and deploy AI models that analyze continuous data generated in the vehicle during the day-to-day operation, including system logs, traces, and vehicle internal signals (Ethernet and CAN) to detect and predict the health of different sub-systems and anticipate failures in real-time. You will own the end-to-end ML pipeline-from data ingestion and model training to deployment on resource-constrained edge devices and model optimization. You will work in a fast-paced startup environment where your code will directly impact fleet reliability and build the next generation of the self-aware vehicle. You will be expected to collaborate with other leading developers who have a deep understanding and expertise of vehicle software and systems, and other AI developers working on MLOps and integration of AI models on vehicles expected to be on the road today. Expect to experiment with cutting-edge model architectures and best-in-class development tools. This is a hybrid role out of our Sunnyvale, CA, where you will be expected to work in our office 3 days a week. Responsibilities: Build and train AI Edge models (e.g., Transformers, LLM, CNN, LSTM, Trees) to process unstructured application logs, kernel traces, and multi-modalities. Integrating ML flows, including cloud-based LLM APIs (Gemini, OpenAI, Claude), with emphasis on synthetic data creation. Develop algorithms to automatically cluster log patterns and detect software regressions, race conditions, or crash precursors. Design unsupervised and supervised learning models (e.g., Autoencoders, Isolation Forests) to monitor time-series data from CAN bus and on-board sensors. Implement logic to correlate signal anomalies (e.g., ADAS drifts, sensor spikes, latency jitters) across different modalities with system events to identify root causes. Port and optimize PyTorch/TensorFlow models into production-grade models for execution on CPU/GPU-bound targets or embedded NPUs. Apply quantization, pruning, distillation, and memory optimization to ensure models run within strict RAM/Flash budgets. Define the data strategy for on-device filtering: pre-processing on device and decide which data is processed locally versus processed in the cloud. Lead the architecture for the edge ML pipeline and mentor junior engineers on best practices for embedded AI. Requirements: Bachelor's degree in Computer Science, Electrical Engineering, Software Engineering, or a related field. 10+ years in Machine Learning Engineering, with 3+ years focused on Edge AI or Embedded Systems. Proven experience mentoring junior engineers in software development. Expert Python (for training) and decent working knowledge of modern C++ (C+/17 for inference). Deep proficiency with PyTorch or TensorFlow, and experience with inference engines like ONNX, TFLite, or TVM. Experience with NLP techniques for textual data parsing, sequence modeling (RNN/GRU), vector store, or lightweight LLMs/SLMs. Experience with libraries like scikit-learn, tslearn, or statsmodels for anomaly detection on sensor data. Proven ability to lead technical projects from concept to production in an ambiguous, fast-paced environment. Ability to communicate with stakeholders and articulate trade-offs. Experience deploying to Edge environments (e.g., ARM-based), managing memory manually, and working with limited compute resources. Candidates with a strong Computer Vision (CV) / ADAS track record are highly encouraged to apply! Desired Skills: MS/PhD in Computer Science, Engineering, or related fields. Familiarity with Edge systems and preferably automotive formats (CAN, DBC, UDS, SOME/IP, or MQTT. Understanding of Linux/QNX kernel logs (dmesg), process states, and OS-level debugging. Experience with NVIDIA TensorRT, Qualcomm SNPE. Sunnyvale HQ Benefits & Perks Offered: Health care plan (Medical, Dental & Vision) Flexible and Dependent Care Expense program Retirement plan (401k) Life Insurance (Basic, Voluntary & AD&D) Unlimited paid time off per year, 14+ paid holidays Hybrid office work arrangement Complimentary lunches, snacks, and beverages during on-site working days Wellness benefit allowance Phone & Internet reimbursement Computer Accessory Allowance The posted salary range is a general guideline and represents a good faith estimate of what Sonatus ("Company") could reasonably expect to pay for a base salary for this position. The pay offered to a selected candidate will be determined based on factors such as (but not limited to) the scope and responsibilities of the position, the qualifications of the selected candidate, departmental budget availability, geographic location and external market pay for comparable jobs. The Company reserves the right to modify this range in the future, as needed, as market conditions change. Base Salary Pay Range $227,000-$300,000 USD
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. Five9 is one of the world's leading cloud contact center platforms. Our AI-powered Automated Quality Management (AQM) product is transforming how enterprises measure, evaluate, and improve agent and interaction quality at scale. As we build out our dedicated AQM Prompt Engineering capability within Specialized Services, we are hiring a Senior WEM AQM Prompt Engineer to lead and shape this practice. This senior role goes beyond hands-on prompt design. You will be the technical authority for Five9 AQM prompt engineering - setting standards, building reusable assets, mentoring the Prompt Engineer team, and acting as the go-to expert for complex customer escalations and pre-sales engagements. If you have spent a significant portion of your career building sophisticated Speech Analytics programs in Verint, NICE, or Genesys - and you have developed serious prompt engineering skills alongside that domain expertise - this role was designed for you. Why Your Speech Analytics Background is the Foundation for This Role Building effective AQM prompts follows the same disciplined, iterative methodology that defines expert-level Speech Analytics category design: Speech Analytics / Verint AQMFive9 AQM Prompt EngineeringDefine category intent & scopeDefine prompt intent & evaluation scopeSelect & tag representative call samplesSelect evaluation call samplesIteratively tune keyword / phrase setsIteratively refine prompt wording & logicTest against live call trafficTest prompt outputs against real callsValidate recall & precision metricsValidate pass / fail accuracy metricsDocument & hand off to QA teamDocument prompts for QA workflow At the senior level, this role adds a leadership and practice-building dimension: you will set the standards the rest of the Prompt Engineering team follows and serve as the technical authority across the most complex customer engagements. Key Responsibilities - Practice Leadership • Serve as the technical lead and subject matter authority for Five9 AQM Prompt Engineering within Specialized Services. • Define and own the prompt engineering methodology, standards, quality gates, and best practice documentation for the team. • Build and maintain a curated, reusable prompt library of validated evaluation templates spanning key verticals and QM use cases. • Mentor and technically guide the WEM AQM Prompt Engineer team (x3), conducting prompt reviews and providing structured feedback. • Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key Responsibilities - Customer Delivery • Lead prompt engineering on the most complex and high-value AQM customer engagements, from initial scoping through production deployment. • Act as the senior escalation point for prompt performance issues, failure analysis, and remediation on live customer deployments. • Partner with customers at senior stakeholder level to translate QM strategy and compliance requirements into precise, measurable AQM evaluation criteria. • Design and execute prompt validation frameworks - including representative call sampling, precision/recall analysis, and A/B prompt testing - to ensure accuracy before production release. Key Responsibilities - Product & Pre-Sales Collaboration • Serve as the Prompt Engineering SME in pre-sales engagements, demonstrating Five9 AQM capabilities and advising prospective customers on prompt-driven evaluation design. • Provide structured product feedback to the Five9 AQM product team based on field experience, identifying gaps, limitations, and enhancement opportunities. • Develop and deliver internal training and enablement for consultants and implementation staff on AQM prompt engineering techniques. • Represent the Prompt Engineering practice in cross-functional forums and contribute to go-to-market materials, customer case studies, and thought leadership content. Key Requirements • Minimum 5 years' hands-on experience as a Speech Analytics Lead, Senior QM Analyst, or WEM Consultant with deep expertise in category/topic design, program governance, and optimization in Verint (Categories/Category Sets), NICE CXne (Category Sets), Genesys (Topics), or equivalent - at a seniority level sufficient to have owned or led an analytics program. • Formal prompt engineering qualification or a verifiable, substantive portfolio of production-grade prompt design work, including evidence of iterative refinement, performance measurement, and documentation. • Demonstrable understanding of LLM behavior - including how prompt structure, instruction framing, persona assignment, chain-of-thought, and few-shot examples affect model outputs. • Strong background in Quality Management design: experience building evaluation frameworks, scorecard design, calibration processes, and QM governance. • Experience leading or mentoring technical team members in a delivery or consulting environment. • Proven track record of translating complex business requirements into structured, testable, and scalable technical solutions. • Excellent written and verbal communication skills; comfortable presenting to senior client stakeholders and internal leadership. • BA/BS or equivalent experience; advanced degree in a relevant discipline is advantageous. Preferred Qualifications • Formal prompt engineering certification: DeepLearning.AI Prompt Engineering for Developers, Anthropic Prompt Engineering, OpenAI Prompt Engineering, DAIR.AI Prompt Engineering Guide, or equivalent industry-recognized program. • Direct experience with one or more leading LLM platforms at an advanced level (OpenAI GPT-4+, Anthropic Claude, Google Gemini, Meta Llama, or similar), including API usage, system prompt design, and evaluation harness development. • Prior hands-on experience with Five9 AQM or another AI-native QM solution (e.g., NICE Enlighten QM, Genesys Cloud AI Quality, Medallia, Qualtrics). • Experience contributing to or owning a Speech Analytics or QM center of excellence - including methodology documentation, internal training, and cross-client best practice standardization. • Familiarity with prompt evaluation frameworks (e.g., LLM-as-judge, human-in-the-loop evaluation pipelines, or equivalent). • Exposure to NLP concepts - tokenization, embeddings, semantic similarity - sufficient to reason about why a prompt succeeds or fails. • Experience with JSON, Python scripting, or no-code automation tools as applied to prompt testing and workflow automation. • Background in contact center compliance-driven QM (financial services, healthcare, utilities) where regulatory accuracy requirements add additional complexity to evaluation design. Key Competencies • Technical Authority - recognized internally and externally as the go-to expert on LLM prompt design for contact Centre quality evaluation. • Analytical Rigor - applies statistical thinking and structured measurement to validate and continuously improve prompt performance. • Linguistic Mastery - exceptional command of written instruction; understands how nuance, framing, and ambiguity in prompt language directly influence AI outputs. • Leadership & Coaching - ability to elevate the capability of others through structured mentoring, peer review, and knowledge transfer. • Strategic Customer Engagement - trusted advisor to senior client stakeholders on the use of AI in quality management. • Innovation Orientation - proactively monitors the LLM landscape and identifies opportunities to improve Five9 AQM prompt design practices. • Operational Excellence - drives the team toward repeatable, scalable processes without sacrificing quality or creativity. About Five9 WEM Specialized Services The Specialized Services team within Five9 Professional Services delivers expert-led implementation, optimization, and advisory engagements across the Five9 WEM portfolio. Our AQM Prompt Engineering practice is a newly established Centre of excellence, designed to ensure that customers realize the full potential of Five9's AI-powered quality management capabilities. As the founding senior hire into this practice, the Senior WEM AQM Prompt Engineer will have a rare opportunity to define how prompt engineering is done at Five9 - and to build something genuinely new. Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally . click apply for full job details
08/05/2026
Full time
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. Five9 is one of the world's leading cloud contact center platforms. Our AI-powered Automated Quality Management (AQM) product is transforming how enterprises measure, evaluate, and improve agent and interaction quality at scale. As we build out our dedicated AQM Prompt Engineering capability within Specialized Services, we are hiring a Senior WEM AQM Prompt Engineer to lead and shape this practice. This senior role goes beyond hands-on prompt design. You will be the technical authority for Five9 AQM prompt engineering - setting standards, building reusable assets, mentoring the Prompt Engineer team, and acting as the go-to expert for complex customer escalations and pre-sales engagements. If you have spent a significant portion of your career building sophisticated Speech Analytics programs in Verint, NICE, or Genesys - and you have developed serious prompt engineering skills alongside that domain expertise - this role was designed for you. Why Your Speech Analytics Background is the Foundation for This Role Building effective AQM prompts follows the same disciplined, iterative methodology that defines expert-level Speech Analytics category design: Speech Analytics / Verint AQMFive9 AQM Prompt EngineeringDefine category intent & scopeDefine prompt intent & evaluation scopeSelect & tag representative call samplesSelect evaluation call samplesIteratively tune keyword / phrase setsIteratively refine prompt wording & logicTest against live call trafficTest prompt outputs against real callsValidate recall & precision metricsValidate pass / fail accuracy metricsDocument & hand off to QA teamDocument prompts for QA workflow At the senior level, this role adds a leadership and practice-building dimension: you will set the standards the rest of the Prompt Engineering team follows and serve as the technical authority across the most complex customer engagements. Key Responsibilities - Practice Leadership • Serve as the technical lead and subject matter authority for Five9 AQM Prompt Engineering within Specialized Services. • Define and own the prompt engineering methodology, standards, quality gates, and best practice documentation for the team. • Build and maintain a curated, reusable prompt library of validated evaluation templates spanning key verticals and QM use cases. • Mentor and technically guide the WEM AQM Prompt Engineer team (x3), conducting prompt reviews and providing structured feedback. • Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key Responsibilities - Customer Delivery • Lead prompt engineering on the most complex and high-value AQM customer engagements, from initial scoping through production deployment. • Act as the senior escalation point for prompt performance issues, failure analysis, and remediation on live customer deployments. • Partner with customers at senior stakeholder level to translate QM strategy and compliance requirements into precise, measurable AQM evaluation criteria. • Design and execute prompt validation frameworks - including representative call sampling, precision/recall analysis, and A/B prompt testing - to ensure accuracy before production release. Key Responsibilities - Product & Pre-Sales Collaboration • Serve as the Prompt Engineering SME in pre-sales engagements, demonstrating Five9 AQM capabilities and advising prospective customers on prompt-driven evaluation design. • Provide structured product feedback to the Five9 AQM product team based on field experience, identifying gaps, limitations, and enhancement opportunities. • Develop and deliver internal training and enablement for consultants and implementation staff on AQM prompt engineering techniques. • Represent the Prompt Engineering practice in cross-functional forums and contribute to go-to-market materials, customer case studies, and thought leadership content. Key Requirements • Minimum 5 years' hands-on experience as a Speech Analytics Lead, Senior QM Analyst, or WEM Consultant with deep expertise in category/topic design, program governance, and optimization in Verint (Categories/Category Sets), NICE CXne (Category Sets), Genesys (Topics), or equivalent - at a seniority level sufficient to have owned or led an analytics program. • Formal prompt engineering qualification or a verifiable, substantive portfolio of production-grade prompt design work, including evidence of iterative refinement, performance measurement, and documentation. • Demonstrable understanding of LLM behavior - including how prompt structure, instruction framing, persona assignment, chain-of-thought, and few-shot examples affect model outputs. • Strong background in Quality Management design: experience building evaluation frameworks, scorecard design, calibration processes, and QM governance. • Experience leading or mentoring technical team members in a delivery or consulting environment. • Proven track record of translating complex business requirements into structured, testable, and scalable technical solutions. • Excellent written and verbal communication skills; comfortable presenting to senior client stakeholders and internal leadership. • BA/BS or equivalent experience; advanced degree in a relevant discipline is advantageous. Preferred Qualifications • Formal prompt engineering certification: DeepLearning.AI Prompt Engineering for Developers, Anthropic Prompt Engineering, OpenAI Prompt Engineering, DAIR.AI Prompt Engineering Guide, or equivalent industry-recognized program. • Direct experience with one or more leading LLM platforms at an advanced level (OpenAI GPT-4+, Anthropic Claude, Google Gemini, Meta Llama, or similar), including API usage, system prompt design, and evaluation harness development. • Prior hands-on experience with Five9 AQM or another AI-native QM solution (e.g., NICE Enlighten QM, Genesys Cloud AI Quality, Medallia, Qualtrics). • Experience contributing to or owning a Speech Analytics or QM center of excellence - including methodology documentation, internal training, and cross-client best practice standardization. • Familiarity with prompt evaluation frameworks (e.g., LLM-as-judge, human-in-the-loop evaluation pipelines, or equivalent). • Exposure to NLP concepts - tokenization, embeddings, semantic similarity - sufficient to reason about why a prompt succeeds or fails. • Experience with JSON, Python scripting, or no-code automation tools as applied to prompt testing and workflow automation. • Background in contact center compliance-driven QM (financial services, healthcare, utilities) where regulatory accuracy requirements add additional complexity to evaluation design. Key Competencies • Technical Authority - recognized internally and externally as the go-to expert on LLM prompt design for contact Centre quality evaluation. • Analytical Rigor - applies statistical thinking and structured measurement to validate and continuously improve prompt performance. • Linguistic Mastery - exceptional command of written instruction; understands how nuance, framing, and ambiguity in prompt language directly influence AI outputs. • Leadership & Coaching - ability to elevate the capability of others through structured mentoring, peer review, and knowledge transfer. • Strategic Customer Engagement - trusted advisor to senior client stakeholders on the use of AI in quality management. • Innovation Orientation - proactively monitors the LLM landscape and identifies opportunities to improve Five9 AQM prompt design practices. • Operational Excellence - drives the team toward repeatable, scalable processes without sacrificing quality or creativity. About Five9 WEM Specialized Services The Specialized Services team within Five9 Professional Services delivers expert-led implementation, optimization, and advisory engagements across the Five9 WEM portfolio. Our AQM Prompt Engineering practice is a newly established Centre of excellence, designed to ensure that customers realize the full potential of Five9's AI-powered quality management capabilities. As the founding senior hire into this practice, the Senior WEM AQM Prompt Engineer will have a rare opportunity to define how prompt engineering is done at Five9 - and to build something genuinely new. Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally . click apply for full job details
Job Description Job Description Zoox is on an ambitious journey to develop a full-stack autonomous mobility solution for cities and safely deploy such a robotaxi solution. The System Design and Mission Assurance (SDMA) team plays a foundational role in the company's success, responsible for constructing the safety case and fail-operational design for our autonomous driving robots before public road deployment. You will be part of an organization with strong leadership and a transparent, respectful culture that enables you to reach your full potential. In this role, you will: Develop, coordinate, and execute verification and validation test plans for Pose monitors using Software-in-the-Loop (SIL) and Hardware-in-the-Loop (HIL) environments. Analyze and triage pipeline results that contribute to key launch-blocking metrics for software and system releases. Collaborate with cross-functional teams, including software developers, hardware engineers, systems engineers, simulation developers, and safety experts, to identify and mitigate risks. Track and report test campaign results to demonstrate feature readiness during safety clearance for each Zoox milestone. Develop metrics and performance benchmarks for V&V campaigns, driving continuous improvement of test coverage and efficiency. Contribute to the development of risk analysis models to quantify behavioral risk when testing gaps exist. Qualifications MS or higher degree in a technical field, such as Mechatronics, Mechanical, Electrical, Aerospace, or Systems Engineering. 5+ years of hands-on systems integration and validation experience (simulation testing, SIL/HIL testing, structured testing, and end-to-end product validation) with a focus on delivering complex systems to production. Proficiency in Python for data analysis, test scripting, and automation. Experience with requirements and test traceability tools such as Polarion, Jama, DOORS, or TestRails. Strong systems-level thinking and the ability to understand and navigate complex technical systems. Basic proficiency with probability and statistics. Ability to grapple with ambiguity and collaborate with cross-functional teams. Bonus Qualifications Experience with autonomous vehicles, robotics, or other safety-critical systems. Experience with simulation environments and hardware-in-the-loop (HIL) testing for pose estimation or perception systems. Experience with fault protection system design and validation. Familiarity with industry safety standards such as ISO 26262, ISO 21448 (SOTIF), ARP 4754, or ARP 4761. Experience with Linux systems and containerized testing environments. Proficiency with SQL for test data analysis. Base Salary Range There are three major components to compensation for this position: salary, Amazon Restricted Stock Units (RSUs), and Zoox Stock Appreciation Rights. A sign-on bonus may be offered as part of the compensation package. The listed range applies only to the base salary. Compensation will vary based on geographic location and level. Leveling, as well as positioning within a level, is determined by a range of factors, including, but not limited to, a candidate's relevant years of experience, domain knowledge, and interview performance. The salary range listed in this posting is representative of the range of levels Zoox is considering for this position. Zoox also offers a comprehensive package of benefits, including paid time off (e.g. sick leave, vacation, bereavement), unpaid time off, Zoox Stock Appreciation Rights, Amazon RSUs, health insurance, long-term care insurance, long-term and short-term disability insurance, and life insurance. About Zoox Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We're looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team. Follow us on LinkedIn Accommodations If you need an accommodation to participate in the application or interview process please reach out to or your assigned recruiter. A Final Note: You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
08/05/2026
Full time
Job Description Job Description Zoox is on an ambitious journey to develop a full-stack autonomous mobility solution for cities and safely deploy such a robotaxi solution. The System Design and Mission Assurance (SDMA) team plays a foundational role in the company's success, responsible for constructing the safety case and fail-operational design for our autonomous driving robots before public road deployment. You will be part of an organization with strong leadership and a transparent, respectful culture that enables you to reach your full potential. In this role, you will: Develop, coordinate, and execute verification and validation test plans for Pose monitors using Software-in-the-Loop (SIL) and Hardware-in-the-Loop (HIL) environments. Analyze and triage pipeline results that contribute to key launch-blocking metrics for software and system releases. Collaborate with cross-functional teams, including software developers, hardware engineers, systems engineers, simulation developers, and safety experts, to identify and mitigate risks. Track and report test campaign results to demonstrate feature readiness during safety clearance for each Zoox milestone. Develop metrics and performance benchmarks for V&V campaigns, driving continuous improvement of test coverage and efficiency. Contribute to the development of risk analysis models to quantify behavioral risk when testing gaps exist. Qualifications MS or higher degree in a technical field, such as Mechatronics, Mechanical, Electrical, Aerospace, or Systems Engineering. 5+ years of hands-on systems integration and validation experience (simulation testing, SIL/HIL testing, structured testing, and end-to-end product validation) with a focus on delivering complex systems to production. Proficiency in Python for data analysis, test scripting, and automation. Experience with requirements and test traceability tools such as Polarion, Jama, DOORS, or TestRails. Strong systems-level thinking and the ability to understand and navigate complex technical systems. Basic proficiency with probability and statistics. Ability to grapple with ambiguity and collaborate with cross-functional teams. Bonus Qualifications Experience with autonomous vehicles, robotics, or other safety-critical systems. Experience with simulation environments and hardware-in-the-loop (HIL) testing for pose estimation or perception systems. Experience with fault protection system design and validation. Familiarity with industry safety standards such as ISO 26262, ISO 21448 (SOTIF), ARP 4754, or ARP 4761. Experience with Linux systems and containerized testing environments. Proficiency with SQL for test data analysis. Base Salary Range There are three major components to compensation for this position: salary, Amazon Restricted Stock Units (RSUs), and Zoox Stock Appreciation Rights. A sign-on bonus may be offered as part of the compensation package. The listed range applies only to the base salary. Compensation will vary based on geographic location and level. Leveling, as well as positioning within a level, is determined by a range of factors, including, but not limited to, a candidate's relevant years of experience, domain knowledge, and interview performance. The salary range listed in this posting is representative of the range of levels Zoox is considering for this position. Zoox also offers a comprehensive package of benefits, including paid time off (e.g. sick leave, vacation, bereavement), unpaid time off, Zoox Stock Appreciation Rights, Amazon RSUs, health insurance, long-term care insurance, long-term and short-term disability insurance, and life insurance. About Zoox Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We're looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team. Follow us on LinkedIn Accommodations If you need an accommodation to participate in the application or interview process please reach out to or your assigned recruiter. A Final Note: You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. Five9 is one of the world's leading cloud contact center platforms. Our AI-powered Automated Quality Management (AQM) product is transforming how enterprises measure, evaluate, and improve agent and interaction quality at scale. As we build out our dedicated AQM Prompt Engineering capability within Specialized Services, we are hiring a Senior WEM AQM Prompt Engineer to lead and shape this practice. This senior role goes beyond hands-on prompt design. You will be the technical authority for Five9 AQM prompt engineering - setting standards, building reusable assets, mentoring the Prompt Engineer team, and acting as the go-to expert for complex customer escalations and pre-sales engagements. If you have spent a significant portion of your career building sophisticated Speech Analytics programs in Verint, NICE, or Genesys - and you have developed serious prompt engineering skills alongside that domain expertise - this role was designed for you. Why Your Speech Analytics Background is the Foundation for This Role Building effective AQM prompts follows the same disciplined, iterative methodology that defines expert-level Speech Analytics category design: Speech Analytics / Verint AQMFive9 AQM Prompt EngineeringDefine category intent & scopeDefine prompt intent & evaluation scopeSelect & tag representative call samplesSelect evaluation call samplesIteratively tune keyword / phrase setsIteratively refine prompt wording & logicTest against live call trafficTest prompt outputs against real callsValidate recall & precision metricsValidate pass / fail accuracy metricsDocument & hand off to QA teamDocument prompts for QA workflow At the senior level, this role adds a leadership and practice-building dimension: you will set the standards the rest of the Prompt Engineering team follows and serve as the technical authority across the most complex customer engagements. Key Responsibilities - Practice Leadership • Serve as the technical lead and subject matter authority for Five9 AQM Prompt Engineering within Specialized Services. • Define and own the prompt engineering methodology, standards, quality gates, and best practice documentation for the team. • Build and maintain a curated, reusable prompt library of validated evaluation templates spanning key verticals and QM use cases. • Mentor and technically guide the WEM AQM Prompt Engineer team (x3), conducting prompt reviews and providing structured feedback. • Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key Responsibilities - Customer Delivery • Lead prompt engineering on the most complex and high-value AQM customer engagements, from initial scoping through production deployment. • Act as the senior escalation point for prompt performance issues, failure analysis, and remediation on live customer deployments. • Partner with customers at senior stakeholder level to translate QM strategy and compliance requirements into precise, measurable AQM evaluation criteria. • Design and execute prompt validation frameworks - including representative call sampling, precision/recall analysis, and A/B prompt testing - to ensure accuracy before production release. Key Responsibilities - Product & Pre-Sales Collaboration • Serve as the Prompt Engineering SME in pre-sales engagements, demonstrating Five9 AQM capabilities and advising prospective customers on prompt-driven evaluation design. • Provide structured product feedback to the Five9 AQM product team based on field experience, identifying gaps, limitations, and enhancement opportunities. • Develop and deliver internal training and enablement for consultants and implementation staff on AQM prompt engineering techniques. • Represent the Prompt Engineering practice in cross-functional forums and contribute to go-to-market materials, customer case studies, and thought leadership content. Key Requirements • Minimum 5 years' hands-on experience as a Speech Analytics Lead, Senior QM Analyst, or WEM Consultant with deep expertise in category/topic design, program governance, and optimization in Verint (Categories/Category Sets), NICE CXne (Category Sets), Genesys (Topics), or equivalent - at a seniority level sufficient to have owned or led an analytics program. • Formal prompt engineering qualification or a verifiable, substantive portfolio of production-grade prompt design work, including evidence of iterative refinement, performance measurement, and documentation. • Demonstrable understanding of LLM behavior - including how prompt structure, instruction framing, persona assignment, chain-of-thought, and few-shot examples affect model outputs. • Strong background in Quality Management design: experience building evaluation frameworks, scorecard design, calibration processes, and QM governance. • Experience leading or mentoring technical team members in a delivery or consulting environment. • Proven track record of translating complex business requirements into structured, testable, and scalable technical solutions. • Excellent written and verbal communication skills; comfortable presenting to senior client stakeholders and internal leadership. • BA/BS or equivalent experience; advanced degree in a relevant discipline is advantageous. Preferred Qualifications • Formal prompt engineering certification: DeepLearning.AI Prompt Engineering for Developers, Anthropic Prompt Engineering, OpenAI Prompt Engineering, DAIR.AI Prompt Engineering Guide, or equivalent industry-recognized program. • Direct experience with one or more leading LLM platforms at an advanced level (OpenAI GPT-4+, Anthropic Claude, Google Gemini, Meta Llama, or similar), including API usage, system prompt design, and evaluation harness development. • Prior hands-on experience with Five9 AQM or another AI-native QM solution (e.g., NICE Enlighten QM, Genesys Cloud AI Quality, Medallia, Qualtrics). • Experience contributing to or owning a Speech Analytics or QM center of excellence - including methodology documentation, internal training, and cross-client best practice standardization. • Familiarity with prompt evaluation frameworks (e.g., LLM-as-judge, human-in-the-loop evaluation pipelines, or equivalent). • Exposure to NLP concepts - tokenization, embeddings, semantic similarity - sufficient to reason about why a prompt succeeds or fails. • Experience with JSON, Python scripting, or no-code automation tools as applied to prompt testing and workflow automation. • Background in contact center compliance-driven QM (financial services, healthcare, utilities) where regulatory accuracy requirements add additional complexity to evaluation design. Key Competencies • Technical Authority - recognized internally and externally as the go-to expert on LLM prompt design for contact Centre quality evaluation. • Analytical Rigor - applies statistical thinking and structured measurement to validate and continuously improve prompt performance. • Linguistic Mastery - exceptional command of written instruction; understands how nuance, framing, and ambiguity in prompt language directly influence AI outputs. • Leadership & Coaching - ability to elevate the capability of others through structured mentoring, peer review, and knowledge transfer. • Strategic Customer Engagement - trusted advisor to senior client stakeholders on the use of AI in quality management. • Innovation Orientation - proactively monitors the LLM landscape and identifies opportunities to improve Five9 AQM prompt design practices. • Operational Excellence - drives the team toward repeatable, scalable processes without sacrificing quality or creativity. About Five9 WEM Specialized Services The Specialized Services team within Five9 Professional Services delivers expert-led implementation, optimization, and advisory engagements across the Five9 WEM portfolio. Our AQM Prompt Engineering practice is a newly established Centre of excellence, designed to ensure that customers realize the full potential of Five9's AI-powered quality management capabilities. As the founding senior hire into this practice, the Senior WEM AQM Prompt Engineer will have a rare opportunity to define how prompt engineering is done at Five9 - and to build something genuinely new. Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally . click apply for full job details
08/05/2026
Full time
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. Five9 is one of the world's leading cloud contact center platforms. Our AI-powered Automated Quality Management (AQM) product is transforming how enterprises measure, evaluate, and improve agent and interaction quality at scale. As we build out our dedicated AQM Prompt Engineering capability within Specialized Services, we are hiring a Senior WEM AQM Prompt Engineer to lead and shape this practice. This senior role goes beyond hands-on prompt design. You will be the technical authority for Five9 AQM prompt engineering - setting standards, building reusable assets, mentoring the Prompt Engineer team, and acting as the go-to expert for complex customer escalations and pre-sales engagements. If you have spent a significant portion of your career building sophisticated Speech Analytics programs in Verint, NICE, or Genesys - and you have developed serious prompt engineering skills alongside that domain expertise - this role was designed for you. Why Your Speech Analytics Background is the Foundation for This Role Building effective AQM prompts follows the same disciplined, iterative methodology that defines expert-level Speech Analytics category design: Speech Analytics / Verint AQMFive9 AQM Prompt EngineeringDefine category intent & scopeDefine prompt intent & evaluation scopeSelect & tag representative call samplesSelect evaluation call samplesIteratively tune keyword / phrase setsIteratively refine prompt wording & logicTest against live call trafficTest prompt outputs against real callsValidate recall & precision metricsValidate pass / fail accuracy metricsDocument & hand off to QA teamDocument prompts for QA workflow At the senior level, this role adds a leadership and practice-building dimension: you will set the standards the rest of the Prompt Engineering team follows and serve as the technical authority across the most complex customer engagements. Key Responsibilities - Practice Leadership • Serve as the technical lead and subject matter authority for Five9 AQM Prompt Engineering within Specialized Services. • Define and own the prompt engineering methodology, standards, quality gates, and best practice documentation for the team. • Build and maintain a curated, reusable prompt library of validated evaluation templates spanning key verticals and QM use cases. • Mentor and technically guide the WEM AQM Prompt Engineer team (x3), conducting prompt reviews and providing structured feedback. • Drive continuous improvement of the team's prompt design process based on performance data, customer outcomes, and LLM platform developments. Key Responsibilities - Customer Delivery • Lead prompt engineering on the most complex and high-value AQM customer engagements, from initial scoping through production deployment. • Act as the senior escalation point for prompt performance issues, failure analysis, and remediation on live customer deployments. • Partner with customers at senior stakeholder level to translate QM strategy and compliance requirements into precise, measurable AQM evaluation criteria. • Design and execute prompt validation frameworks - including representative call sampling, precision/recall analysis, and A/B prompt testing - to ensure accuracy before production release. Key Responsibilities - Product & Pre-Sales Collaboration • Serve as the Prompt Engineering SME in pre-sales engagements, demonstrating Five9 AQM capabilities and advising prospective customers on prompt-driven evaluation design. • Provide structured product feedback to the Five9 AQM product team based on field experience, identifying gaps, limitations, and enhancement opportunities. • Develop and deliver internal training and enablement for consultants and implementation staff on AQM prompt engineering techniques. • Represent the Prompt Engineering practice in cross-functional forums and contribute to go-to-market materials, customer case studies, and thought leadership content. Key Requirements • Minimum 5 years' hands-on experience as a Speech Analytics Lead, Senior QM Analyst, or WEM Consultant with deep expertise in category/topic design, program governance, and optimization in Verint (Categories/Category Sets), NICE CXne (Category Sets), Genesys (Topics), or equivalent - at a seniority level sufficient to have owned or led an analytics program. • Formal prompt engineering qualification or a verifiable, substantive portfolio of production-grade prompt design work, including evidence of iterative refinement, performance measurement, and documentation. • Demonstrable understanding of LLM behavior - including how prompt structure, instruction framing, persona assignment, chain-of-thought, and few-shot examples affect model outputs. • Strong background in Quality Management design: experience building evaluation frameworks, scorecard design, calibration processes, and QM governance. • Experience leading or mentoring technical team members in a delivery or consulting environment. • Proven track record of translating complex business requirements into structured, testable, and scalable technical solutions. • Excellent written and verbal communication skills; comfortable presenting to senior client stakeholders and internal leadership. • BA/BS or equivalent experience; advanced degree in a relevant discipline is advantageous. Preferred Qualifications • Formal prompt engineering certification: DeepLearning.AI Prompt Engineering for Developers, Anthropic Prompt Engineering, OpenAI Prompt Engineering, DAIR.AI Prompt Engineering Guide, or equivalent industry-recognized program. • Direct experience with one or more leading LLM platforms at an advanced level (OpenAI GPT-4+, Anthropic Claude, Google Gemini, Meta Llama, or similar), including API usage, system prompt design, and evaluation harness development. • Prior hands-on experience with Five9 AQM or another AI-native QM solution (e.g., NICE Enlighten QM, Genesys Cloud AI Quality, Medallia, Qualtrics). • Experience contributing to or owning a Speech Analytics or QM center of excellence - including methodology documentation, internal training, and cross-client best practice standardization. • Familiarity with prompt evaluation frameworks (e.g., LLM-as-judge, human-in-the-loop evaluation pipelines, or equivalent). • Exposure to NLP concepts - tokenization, embeddings, semantic similarity - sufficient to reason about why a prompt succeeds or fails. • Experience with JSON, Python scripting, or no-code automation tools as applied to prompt testing and workflow automation. • Background in contact center compliance-driven QM (financial services, healthcare, utilities) where regulatory accuracy requirements add additional complexity to evaluation design. Key Competencies • Technical Authority - recognized internally and externally as the go-to expert on LLM prompt design for contact Centre quality evaluation. • Analytical Rigor - applies statistical thinking and structured measurement to validate and continuously improve prompt performance. • Linguistic Mastery - exceptional command of written instruction; understands how nuance, framing, and ambiguity in prompt language directly influence AI outputs. • Leadership & Coaching - ability to elevate the capability of others through structured mentoring, peer review, and knowledge transfer. • Strategic Customer Engagement - trusted advisor to senior client stakeholders on the use of AI in quality management. • Innovation Orientation - proactively monitors the LLM landscape and identifies opportunities to improve Five9 AQM prompt design practices. • Operational Excellence - drives the team toward repeatable, scalable processes without sacrificing quality or creativity. About Five9 WEM Specialized Services The Specialized Services team within Five9 Professional Services delivers expert-led implementation, optimization, and advisory engagements across the Five9 WEM portfolio. Our AQM Prompt Engineering practice is a newly established Centre of excellence, designed to ensure that customers realize the full potential of Five9's AI-powered quality management capabilities. As the founding senior hire into this practice, the Senior WEM AQM Prompt Engineer will have a rare opportunity to define how prompt engineering is done at Five9 - and to build something genuinely new. Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally . click apply for full job details
Job Description Job Description About Us: We are building a robust, scalable trading platform to serve high-traffic, latency-sensitive applications. Our infrastructure leverages state-of-the-art technologies to support real-time trading while providing unparalleled reliability and performance. Join us to shape the future of our platform and engineering culture. Job Summary: We are looking for a Senior DevOps & Platform Engineer to lead the design, implementation, and management of our AWS-centric infrastructure. You will play a pivotal role in maximizing the velocity of our product engineering team, ensuring platform scalability, reliability, and security. This is a high-impact role, combining elements of DevOps, Platform Engineering, and Site Reliability Engineering (SRE). You will champion best practices, shape the engineering culture, and ensure our platform is robust, efficient, and ready for the future. Key Responsibilities: Platform Engineering Infrastructure Design: Architect and implement scalable infrastructure to support the deployment and management of our trading platform. Developer Tooling: Build and maintain internal tools to streamline developer workflows, including advanced CI/CD pipelines. Infrastructure as Code (IaC): Champion IaC practices using Terraform, CloudFormation, or Pulumi. Core Services Management: Manage and optimize platform-critical services such as: NATS Cluster RabbitMQ AWS RDS PostgreSQL Redis Cluster DevOps Automation and CI/CD: Automate and optimize deployment processes to ensure seamless continuous integration and delivery. Container Orchestration: Manage and scale containerized workloads using Kubernetes and Docker. Cloud Optimization: Monitor and optimize cloud resource usage for performance and cost efficiency. Site Reliability Engineering (SRE) Reliability Metrics: Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Monitoring & Observability: Implement observability tools and dashboards (e.g., Prometheus, Datadog, Grafana) for real-time system monitoring. Incident Management: Lead incident response efforts, conduct root cause analysis, and implement actionable postmortem reviews. Infrastructure Management AWS Expertise: Architect and manage cloud-based systems to handle high-traffic, latency-sensitive applications. Disaster Recovery: Implement robust disaster recovery and business continuity strategies, including backups and multi-region failover. Security Practices: Collaborate with security teams to enforce best practices for IAM, encryption, and compliance. Collaboration & Leadership Cross-Team Collaboration: Partner with software engineers to design infrastructure solutions tailored to their application needs. Culture Building: Help shape the engineering culture, promoting a philosophy of security, velocity, and reliability. Mentorship: Mentor junior engineers and document best practices to drive knowledge sharing and operational excellence. Long-Term Tech Evolution Backend Transition: Contribute to evolving our backend microservices (currently NodeJS, with some Python and C#) towards Go and Rust. Third-Party Integration: Evaluate and integrate critical third-party software and infrastructure, such as payment gateways and mobility stacks. Your Impact: Simplify infrastructure concerns for product teams to accelerate builds, deployments, and scaling. Advocate for modern practices like Zero Trust Networking and continuously improve platform architecture. Balance the demands of product velocity with a well-managed, secure, and scalable platform. Required Skills & Experience: Technical Expertise Cloud Experience: 5-8+ years of hands-on experience with cloud platforms, particularly AWS, including services like EC2, RDS, S3, Lambda, and VPC. Containerization: Proficiency with Docker and Kubernetes (EKS) or ECS. Infrastructure as Code (IaC): Strong experience with Terraform, CloudFormation, or Pulumi. Programming Skills: Proficiency in at least one programming language (e.g., Python, Go, TypeScript/JavaScript, Ruby, Java). DevOps & SRE CI/CD Pipelines: Expertise in building and maintaining CI/CD workflows using tools like GitLab CI, Jenkins, or GitHub Actions. Monitoring Tools: Experience with observability platforms (e.g., Prometheus, Datadog, Grafana). Incident Management: Proven ability to handle incident response, root cause analysis, and postmortem reviews. Soft Skills Problem-Solving: Ability to research, design, and deliver solutions to complex infrastructure challenges. Collaboration: Experience working directly with product engineers to improve workflows incrementally. Leadership: Ownership mindset with the ability to mentor team members and advocate for best practices. Preferred Skills (Nice-to-Have): Familiarity with backend languages like Go or Rust. AWS certifications (e.g., Solutions Architect, DevOps Engineer). Experience with networking concepts (e.g., load balancers, DNS, VPNs) and traffic optimization. Knowledge of emerging CNCF technologies and CI/CD trends. What We Offer: Competitive salary with future equity options Opportunities to work with cutting-edge technologies and evolve our platform. Flexible working hours and a remote-friendly environment. Professional growth through certifications, conferences, and internal training. Collaborative culture focused on innovation and operational excellence.
08/05/2026
Full time
Job Description Job Description About Us: We are building a robust, scalable trading platform to serve high-traffic, latency-sensitive applications. Our infrastructure leverages state-of-the-art technologies to support real-time trading while providing unparalleled reliability and performance. Join us to shape the future of our platform and engineering culture. Job Summary: We are looking for a Senior DevOps & Platform Engineer to lead the design, implementation, and management of our AWS-centric infrastructure. You will play a pivotal role in maximizing the velocity of our product engineering team, ensuring platform scalability, reliability, and security. This is a high-impact role, combining elements of DevOps, Platform Engineering, and Site Reliability Engineering (SRE). You will champion best practices, shape the engineering culture, and ensure our platform is robust, efficient, and ready for the future. Key Responsibilities: Platform Engineering Infrastructure Design: Architect and implement scalable infrastructure to support the deployment and management of our trading platform. Developer Tooling: Build and maintain internal tools to streamline developer workflows, including advanced CI/CD pipelines. Infrastructure as Code (IaC): Champion IaC practices using Terraform, CloudFormation, or Pulumi. Core Services Management: Manage and optimize platform-critical services such as: NATS Cluster RabbitMQ AWS RDS PostgreSQL Redis Cluster DevOps Automation and CI/CD: Automate and optimize deployment processes to ensure seamless continuous integration and delivery. Container Orchestration: Manage and scale containerized workloads using Kubernetes and Docker. Cloud Optimization: Monitor and optimize cloud resource usage for performance and cost efficiency. Site Reliability Engineering (SRE) Reliability Metrics: Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Monitoring & Observability: Implement observability tools and dashboards (e.g., Prometheus, Datadog, Grafana) for real-time system monitoring. Incident Management: Lead incident response efforts, conduct root cause analysis, and implement actionable postmortem reviews. Infrastructure Management AWS Expertise: Architect and manage cloud-based systems to handle high-traffic, latency-sensitive applications. Disaster Recovery: Implement robust disaster recovery and business continuity strategies, including backups and multi-region failover. Security Practices: Collaborate with security teams to enforce best practices for IAM, encryption, and compliance. Collaboration & Leadership Cross-Team Collaboration: Partner with software engineers to design infrastructure solutions tailored to their application needs. Culture Building: Help shape the engineering culture, promoting a philosophy of security, velocity, and reliability. Mentorship: Mentor junior engineers and document best practices to drive knowledge sharing and operational excellence. Long-Term Tech Evolution Backend Transition: Contribute to evolving our backend microservices (currently NodeJS, with some Python and C#) towards Go and Rust. Third-Party Integration: Evaluate and integrate critical third-party software and infrastructure, such as payment gateways and mobility stacks. Your Impact: Simplify infrastructure concerns for product teams to accelerate builds, deployments, and scaling. Advocate for modern practices like Zero Trust Networking and continuously improve platform architecture. Balance the demands of product velocity with a well-managed, secure, and scalable platform. Required Skills & Experience: Technical Expertise Cloud Experience: 5-8+ years of hands-on experience with cloud platforms, particularly AWS, including services like EC2, RDS, S3, Lambda, and VPC. Containerization: Proficiency with Docker and Kubernetes (EKS) or ECS. Infrastructure as Code (IaC): Strong experience with Terraform, CloudFormation, or Pulumi. Programming Skills: Proficiency in at least one programming language (e.g., Python, Go, TypeScript/JavaScript, Ruby, Java). DevOps & SRE CI/CD Pipelines: Expertise in building and maintaining CI/CD workflows using tools like GitLab CI, Jenkins, or GitHub Actions. Monitoring Tools: Experience with observability platforms (e.g., Prometheus, Datadog, Grafana). Incident Management: Proven ability to handle incident response, root cause analysis, and postmortem reviews. Soft Skills Problem-Solving: Ability to research, design, and deliver solutions to complex infrastructure challenges. Collaboration: Experience working directly with product engineers to improve workflows incrementally. Leadership: Ownership mindset with the ability to mentor team members and advocate for best practices. Preferred Skills (Nice-to-Have): Familiarity with backend languages like Go or Rust. AWS certifications (e.g., Solutions Architect, DevOps Engineer). Experience with networking concepts (e.g., load balancers, DNS, VPNs) and traffic optimization. Knowledge of emerging CNCF technologies and CI/CD trends. What We Offer: Competitive salary with future equity options Opportunities to work with cutting-edge technologies and evolve our platform. Flexible working hours and a remote-friendly environment. Professional growth through certifications, conferences, and internal training. Collaborative culture focused on innovation and operational excellence.
Job Description Job Description Grey Matters Defense Solutions specializes in software development and solutions, specialized data analytics algorithms, and innovative remote sensing technologies and operations, with senior-level personnel formerly from DIA, NRO, DARPA, and the US Armed Forces. Grey Matters Defense Solutions differentiates itself by combining its subject matter experts, analysts, software engineers, and its data scientists to create unique artificial intelligence algorithms and various applications. Grey Matters Defense Solutions is seeking a talented and dedicated, Mission Engineer About the job: Grey Matters Defense Solutions is looking for experienced multi-INT analysts/developers to join our Sensor to Shooter Mission Engineering team to help solve mission-focused operational problems. Stationed at customer locations, the Mission Engineering team provides support for local operational needs while also helping integrate with Intel Community and DoD tools, services, and data. Key Responsibilities: The Mission Engineering team is responsible for creating and integrating capabilities into Sensor to Shooter workflows, capturing customer needs, and providing onsite expertise through development of scripts, models, automations, and other creative techniques to overcome mission-limiting capability gaps. Ideal additions to this team will be broadly experienced across data types and intelligence disciplines and are highly motivated problem-solvers. New Mission Engineers must have exceptional communication skills and function effectively within a team as well as independently. Dedication to the mission and an ability to generate creative solutions to hard mission problems is a must. Willingness to travel in support of exercises, surge support, etc. as needed to support Customer requirements. About you: An active Top Secret security clearance 8+ years of IC/DoD analyst experience 5+ years of mission capability experience 5+ years of experience with common DoD analysis and visualization applications Demonstrated ability to communicate with peers, leadership, and customers Excellent analytical skills and problem-solving skills High level of self-initiative and self-motivation with the ability to work under minimal supervision Ability to work effectively in small team settings to solve complex problems Demonstrated ability to bring new or innovative solutions to problems Preferred Skills: Coding experience in a modern software language such as Python, R, or JAVA Broad community knowledge of sensors and disciplines and the ability to leverage and communicate this intelligence, surveillance, and reconnaissance experience to teammates and customers Experience in processing and visualizing data and an understanding of the processes and procedures necessary to request, coordinate, and task collection Knowledge of customer tasking, collection, and processing tools to perform limited analytics and development Understanding of Agile development practices and procedures with the ability to act as a product owner for established needs Salary Range: $100,000 - $160,000 Grey Matters Defense Solutions offer a comprehensive benefits package including medical, dental, vision, life insurance, short-term and long-term disability. Additional Benefits: SEP IRA 25% of base salary PTO Six weeks IBA 12.5% Employee assistance program Employee discount Flexible spending account Health savings account Referral program Grey Matters Defense Solutions' most valuable assets are the more than 60+ employees, consisting of data scientists, custom software developers, and analysts/subject matter experts, with senior-level personnel formerly from DIA, NRO, NSA and the US Armed Forces. Our employees have a depth of analytical knowledge which provides them with deep understanding of managing and delivering products within government systems. Grey Matters Defense Solutions provides transformational leadership building aware-winning teams and products. Join our team of exceptional developers, architects and data scientists! All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran. Powered by JazzHR 7VTpjIgd2h
08/05/2026
Full time
Job Description Job Description Grey Matters Defense Solutions specializes in software development and solutions, specialized data analytics algorithms, and innovative remote sensing technologies and operations, with senior-level personnel formerly from DIA, NRO, DARPA, and the US Armed Forces. Grey Matters Defense Solutions differentiates itself by combining its subject matter experts, analysts, software engineers, and its data scientists to create unique artificial intelligence algorithms and various applications. Grey Matters Defense Solutions is seeking a talented and dedicated, Mission Engineer About the job: Grey Matters Defense Solutions is looking for experienced multi-INT analysts/developers to join our Sensor to Shooter Mission Engineering team to help solve mission-focused operational problems. Stationed at customer locations, the Mission Engineering team provides support for local operational needs while also helping integrate with Intel Community and DoD tools, services, and data. Key Responsibilities: The Mission Engineering team is responsible for creating and integrating capabilities into Sensor to Shooter workflows, capturing customer needs, and providing onsite expertise through development of scripts, models, automations, and other creative techniques to overcome mission-limiting capability gaps. Ideal additions to this team will be broadly experienced across data types and intelligence disciplines and are highly motivated problem-solvers. New Mission Engineers must have exceptional communication skills and function effectively within a team as well as independently. Dedication to the mission and an ability to generate creative solutions to hard mission problems is a must. Willingness to travel in support of exercises, surge support, etc. as needed to support Customer requirements. About you: An active Top Secret security clearance 8+ years of IC/DoD analyst experience 5+ years of mission capability experience 5+ years of experience with common DoD analysis and visualization applications Demonstrated ability to communicate with peers, leadership, and customers Excellent analytical skills and problem-solving skills High level of self-initiative and self-motivation with the ability to work under minimal supervision Ability to work effectively in small team settings to solve complex problems Demonstrated ability to bring new or innovative solutions to problems Preferred Skills: Coding experience in a modern software language such as Python, R, or JAVA Broad community knowledge of sensors and disciplines and the ability to leverage and communicate this intelligence, surveillance, and reconnaissance experience to teammates and customers Experience in processing and visualizing data and an understanding of the processes and procedures necessary to request, coordinate, and task collection Knowledge of customer tasking, collection, and processing tools to perform limited analytics and development Understanding of Agile development practices and procedures with the ability to act as a product owner for established needs Salary Range: $100,000 - $160,000 Grey Matters Defense Solutions offer a comprehensive benefits package including medical, dental, vision, life insurance, short-term and long-term disability. Additional Benefits: SEP IRA 25% of base salary PTO Six weeks IBA 12.5% Employee assistance program Employee discount Flexible spending account Health savings account Referral program Grey Matters Defense Solutions' most valuable assets are the more than 60+ employees, consisting of data scientists, custom software developers, and analysts/subject matter experts, with senior-level personnel formerly from DIA, NRO, NSA and the US Armed Forces. Our employees have a depth of analytical knowledge which provides them with deep understanding of managing and delivering products within government systems. Grey Matters Defense Solutions provides transformational leadership building aware-winning teams and products. Join our team of exceptional developers, architects and data scientists! All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran. Powered by JazzHR 7VTpjIgd2h
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer - Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices. This is a deeply technical role requiring expert-level understanding of compute hardware management, Kubernetes, OpenStack, hypervisors and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications. Key Responsibilities Lead the architecture and design of enterprise compute and hypervisor platform solutions across hardware, OS, virtualization, cloud orchestration, and container orchestration layers Define standards and automation frameworks for bare metal provisioning and lifecycle management Design and implement Bare Metal as a Service (BMaaS) capabilities for scalable infrastructure consumption Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD) Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester Design and maintain PXE-based provisioning environments leveraging Redfish APIs for large-scale server deployments Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation. Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback. Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation Evaluate and standardize enterprise hardware platforms to meet performance, scalability, and reliability requirements Produce detailed high-level and low-level design documentation , build guides, and operational handoff materials Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-ready Participate in on-call escalation support for complex platform-related issues Collaborate globally on change management , documentation, and operational best practices Minimum Qualifications 6 + years of experience in infrastructure engineering, platform engineering, or DevOps with a strong focus on Compute system design Proven experience designing and automating bare metal compute environments at scale Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms Practical experience using Redfish APIs for hardware provisioning, power management, and remote lifecycle operations Deep expertise with Ubuntu Linux in enterprise environments Strong Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Experience designing and deploying production-grade Kubernetes clusters Strong background with enterprise compute hardware platforms , including Cisco UCS, Dell PowerEdge, Supermicro systems & HPE Proficiency with Infrastructure as Code tools (e.g., Terraform, Ansible, or similar) Experience building or supporting CI/CD pipelines for infrastructure and platform automation Strong scripting skills in Python, Bash, or similar languages Demonstrated ability to produce clear, structured technical design documentation Excellent written and verbal communication skills Bachelor's degree in computer science or equivalent professional experience Preferred Qualifications OpenStack, Ubuntu KVM administration. BareMetal as a Service (PXE, Redfish). Kubernetes on BareMetal CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. Background in telco, edge cloud, or large enterprise environments. Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) Master's degree in computer science, IT, Engineering, or a related field preferred; equivalent experience and relevant industry certifications will also be considered What You'll Get A collaborative team that's deeply invested in infrastructure excellence. Complex technical challenges that require creative, scalable solutions. The opportunity to shape a next-generation private cloud platform-built reliability Access to the latest tools, frameworks, and upstream project developments Skills and Attributes: Analytical Thinking & Problem Solving: Demonstrated ability to translate complex, cross-domain requirements into scalable and resilient cloud infrastructure and automation solutions Collaboration & Teamwork: Strong interpersonal and communication skills with a proven track record of effective collaboration across multidisciplinary teams, including developers, operations, security, and product stakeholders Mentorship & Leadership: Passionate about knowledge-sharing and mentorship, with experience guiding junior engineers and fostering a team culture of continuous learning, innovation, and technical excellence in cloud engineering and DevOps practices Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally, the total compensation package for this position may also include an annual performance bonus, stock, and/or other applicable incentive compensation plans. Our total reward package also includes: Health, dental, and vision coverage, beginning on the first day of employment. Five9 covers 100% of the employee portion of the health, dental and vision coverage and shares a high portion of the dependent cost. We also offer Short & Long-Term Disability, Basic Life Insurance, and a 401k saving plan with employer matching. Access to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for all covered employees and their covered dependents. Generous employee stock purchase plan. Paid Time Off, Company paid holidays, paid volunteer hours and 12 weeks paid parental leave. All compensation and benefits are subject to the requirements and restrictions set forth in the applicable plan documents and any written agreements between the parties. The US base salary range for this role is below. $82,300-$228,800 USD Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better we are. Five9 is an equal opportunity employer. View our privacy policy, including our privacy notice to California residents here: -pt/legal. Note: Five9 will never request that an applicant send money as a prerequisite for commencing employment with Five9.
08/05/2026
Full time
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer - Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices. This is a deeply technical role requiring expert-level understanding of compute hardware management, Kubernetes, OpenStack, hypervisors and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications. Key Responsibilities Lead the architecture and design of enterprise compute and hypervisor platform solutions across hardware, OS, virtualization, cloud orchestration, and container orchestration layers Define standards and automation frameworks for bare metal provisioning and lifecycle management Design and implement Bare Metal as a Service (BMaaS) capabilities for scalable infrastructure consumption Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD) Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester Design and maintain PXE-based provisioning environments leveraging Redfish APIs for large-scale server deployments Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation. Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback. Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation Evaluate and standardize enterprise hardware platforms to meet performance, scalability, and reliability requirements Produce detailed high-level and low-level design documentation , build guides, and operational handoff materials Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-ready Participate in on-call escalation support for complex platform-related issues Collaborate globally on change management , documentation, and operational best practices Minimum Qualifications 6 + years of experience in infrastructure engineering, platform engineering, or DevOps with a strong focus on Compute system design Proven experience designing and automating bare metal compute environments at scale Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms Practical experience using Redfish APIs for hardware provisioning, power management, and remote lifecycle operations Deep expertise with Ubuntu Linux in enterprise environments Strong Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Experience designing and deploying production-grade Kubernetes clusters Strong background with enterprise compute hardware platforms , including Cisco UCS, Dell PowerEdge, Supermicro systems & HPE Proficiency with Infrastructure as Code tools (e.g., Terraform, Ansible, or similar) Experience building or supporting CI/CD pipelines for infrastructure and platform automation Strong scripting skills in Python, Bash, or similar languages Demonstrated ability to produce clear, structured technical design documentation Excellent written and verbal communication skills Bachelor's degree in computer science or equivalent professional experience Preferred Qualifications OpenStack, Ubuntu KVM administration. BareMetal as a Service (PXE, Redfish). Kubernetes on BareMetal CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. Background in telco, edge cloud, or large enterprise environments. Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) Master's degree in computer science, IT, Engineering, or a related field preferred; equivalent experience and relevant industry certifications will also be considered What You'll Get A collaborative team that's deeply invested in infrastructure excellence. Complex technical challenges that require creative, scalable solutions. The opportunity to shape a next-generation private cloud platform-built reliability Access to the latest tools, frameworks, and upstream project developments Skills and Attributes: Analytical Thinking & Problem Solving: Demonstrated ability to translate complex, cross-domain requirements into scalable and resilient cloud infrastructure and automation solutions Collaboration & Teamwork: Strong interpersonal and communication skills with a proven track record of effective collaboration across multidisciplinary teams, including developers, operations, security, and product stakeholders Mentorship & Leadership: Passionate about knowledge-sharing and mentorship, with experience guiding junior engineers and fostering a team culture of continuous learning, innovation, and technical excellence in cloud engineering and DevOps practices Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally, the total compensation package for this position may also include an annual performance bonus, stock, and/or other applicable incentive compensation plans. Our total reward package also includes: Health, dental, and vision coverage, beginning on the first day of employment. Five9 covers 100% of the employee portion of the health, dental and vision coverage and shares a high portion of the dependent cost. We also offer Short & Long-Term Disability, Basic Life Insurance, and a 401k saving plan with employer matching. Access to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for all covered employees and their covered dependents. Generous employee stock purchase plan. Paid Time Off, Company paid holidays, paid volunteer hours and 12 weeks paid parental leave. All compensation and benefits are subject to the requirements and restrictions set forth in the applicable plan documents and any written agreements between the parties. The US base salary range for this role is below. $82,300-$228,800 USD Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better we are. Five9 is an equal opportunity employer. View our privacy policy, including our privacy notice to California residents here: -pt/legal. Note: Five9 will never request that an applicant send money as a prerequisite for commencing employment with Five9.
Job Description Job Description Senior Software Engineer, Reliability Remote, US About Nametag Nametag is building the future of secure digital identity. Our mission is to make it easy for people and organizations to prove who they are online, safely and seamlessly. We're pioneering next-generation identity verification and account protection so that users can control their own identity, and companies can build trust without friction. By enabling trust at scale, we help enterprises reduce fraud, streamline onboarding, and deliver secure user experiences that customers love. The Role We're looking for a Senior Software Engineer to own infrastructure, reliability, and platform engineering at Nametag. You'll design and scale the systems that keep our identity verification platform fast, secure, and always available, and build the tooling that makes our engineering team more productive and autonomous. This is a hands-on, high-impact role. You'll work closely with engineering and product leadership to make critical decisions about how we build, deploy, and operate our systems, and then execute on them. If you've thrived in fast-moving environments and care deeply about uptime, correctness, and developer experience, we'd love to meet you. What You'll Do Infrastructure & Reliability Design, build, and maintain scalable, cost-effective cloud infrastructure across AWS and Fly.io. Own deployment strategy and service design across our monolith and microservices, including schema migrations and rollout safety. Manage infrastructure-as-code (CloudFormation, Terraform or equivalent) for reproducible, auditable infrastructure. Identify reliability risks, performance bottlenecks, and security gaps, and address them proactively. Observability & Incident Response Evolve our observability stack: logging, distributed tracing, metrics, and uptime monitoring. Evolve on-call practices and incident response processes that keep us ahead of customer-impacting issues. Champion a culture of reliability: postmortems, runbooks, and continuous improvement after incidents. Developer Enablement Manage and improve our CI/CD pipelines (GitHub Actions) and establish deployment best practices. Build internal platform tooling and abstractions that reduce toil and increase engineering velocity. Partner with product engineers to make infrastructure easy to use correctly and hard to use incorrectly. Data & ML Infrastructure Design and operate data pipelines that support our ML-powered verification systems. Evolve our MLOps infrastructure so models can be trained, evaluated, and deployed safely and repeatedly. Collaboration & Leadership Work closely with engineering and product leadership on technical roadmap decisions. Review code, mentor peers, and help raise the bar on security, reliability, and operational discipline. Communicate infrastructure tradeoffs clearly across technical and non-technical stakeholders. Ideal Qualifications Work Authorization (Required): Applicants must be legally authorized to work in the United States for any employer without current or future need for visa sponsorship. This is a firm requirement. The technical qualifications below are guidelines. We know that no candidate will perfectly match every requirement, and that's okay. If you're passionate about what we're building and have most of the skills below, we'd love to hear from you. Cloud & Infrastructure: Hands-on experience managing and securing cloud infrastructure. AWS required; Fly.io or similar a plus. Infrastructure-as-Code: Production experience with Terraform or equivalent tools. Databases: Deep experience with PostgreSQL, including schema design, migrations, and query performance tuning. CI/CD: Strong experience designing and managing pipelines, GitHub Actions preferred. Languages: Proficiency in modern, type-safe languages. Go strongly preferred. Observability: Experience building and operating logging, tracing, and metrics systems in production. MLOps & Data Pipelines: Prior experience shipping ML infrastructure into production, not just experimentation. Startup experience: You've worked at an early-stage company and know what it means to move fast without compromising the things that matter. Security: Security-minded by default. You design with the threat model in mind, not as an afterthought. What We Value Intellectual horsepower. Quickly grasping complex technical and business concepts. Kindness and integrity. Earning trust is central to how we build relationships with customers and colleagues. Bias for action. We move quickly to deliver impact and protect our customers against fast-moving threats. Ownership mindset. You care about uptime and stability the way a founder cares about the product. Compensation The base salary range for this full-time position is $120,000 to $190,000, plus equity and benefits. Nametag is a founding member of the Open Imperative, publicly committed to pay equity in the technology industry. We post positions with ranges to encourage people of different backgrounds and experiences to apply. Every offer is benchmarked against market data to ensure fairness and consistency. Final compensation is determined by role, level, and additional factors such as skills, experience, and education. Your recruiter or hiring manager can share more details during the hiring process. Culture & Perks At Nametag, we believe trust starts with how we treat each other. We're a remote-first team that values autonomy, inclusivity, and collaboration, with regular in-person time to stay connected and innovate together. Remote-first: Work from anywhere in the US. Our team spans Seattle, San Francisco, Ann Arbor, Denver, New York City, and beyond. Off-sites: We bring the team together once per quarter for in-person collaboration, often off-site in new places. Flexible schedules: Work in your own time zone; we align key meetings across a shared window. We Offer Competitive salary Meaningful equity ownership Comprehensive health benefits (medical, dental, vision) Flexible paid time off Quarterly team off-sites and travel support New computer hardware and equipment An inclusive environment where your voice has impact and your work drives change
08/05/2026
Full time
Job Description Job Description Senior Software Engineer, Reliability Remote, US About Nametag Nametag is building the future of secure digital identity. Our mission is to make it easy for people and organizations to prove who they are online, safely and seamlessly. We're pioneering next-generation identity verification and account protection so that users can control their own identity, and companies can build trust without friction. By enabling trust at scale, we help enterprises reduce fraud, streamline onboarding, and deliver secure user experiences that customers love. The Role We're looking for a Senior Software Engineer to own infrastructure, reliability, and platform engineering at Nametag. You'll design and scale the systems that keep our identity verification platform fast, secure, and always available, and build the tooling that makes our engineering team more productive and autonomous. This is a hands-on, high-impact role. You'll work closely with engineering and product leadership to make critical decisions about how we build, deploy, and operate our systems, and then execute on them. If you've thrived in fast-moving environments and care deeply about uptime, correctness, and developer experience, we'd love to meet you. What You'll Do Infrastructure & Reliability Design, build, and maintain scalable, cost-effective cloud infrastructure across AWS and Fly.io. Own deployment strategy and service design across our monolith and microservices, including schema migrations and rollout safety. Manage infrastructure-as-code (CloudFormation, Terraform or equivalent) for reproducible, auditable infrastructure. Identify reliability risks, performance bottlenecks, and security gaps, and address them proactively. Observability & Incident Response Evolve our observability stack: logging, distributed tracing, metrics, and uptime monitoring. Evolve on-call practices and incident response processes that keep us ahead of customer-impacting issues. Champion a culture of reliability: postmortems, runbooks, and continuous improvement after incidents. Developer Enablement Manage and improve our CI/CD pipelines (GitHub Actions) and establish deployment best practices. Build internal platform tooling and abstractions that reduce toil and increase engineering velocity. Partner with product engineers to make infrastructure easy to use correctly and hard to use incorrectly. Data & ML Infrastructure Design and operate data pipelines that support our ML-powered verification systems. Evolve our MLOps infrastructure so models can be trained, evaluated, and deployed safely and repeatedly. Collaboration & Leadership Work closely with engineering and product leadership on technical roadmap decisions. Review code, mentor peers, and help raise the bar on security, reliability, and operational discipline. Communicate infrastructure tradeoffs clearly across technical and non-technical stakeholders. Ideal Qualifications Work Authorization (Required): Applicants must be legally authorized to work in the United States for any employer without current or future need for visa sponsorship. This is a firm requirement. The technical qualifications below are guidelines. We know that no candidate will perfectly match every requirement, and that's okay. If you're passionate about what we're building and have most of the skills below, we'd love to hear from you. Cloud & Infrastructure: Hands-on experience managing and securing cloud infrastructure. AWS required; Fly.io or similar a plus. Infrastructure-as-Code: Production experience with Terraform or equivalent tools. Databases: Deep experience with PostgreSQL, including schema design, migrations, and query performance tuning. CI/CD: Strong experience designing and managing pipelines, GitHub Actions preferred. Languages: Proficiency in modern, type-safe languages. Go strongly preferred. Observability: Experience building and operating logging, tracing, and metrics systems in production. MLOps & Data Pipelines: Prior experience shipping ML infrastructure into production, not just experimentation. Startup experience: You've worked at an early-stage company and know what it means to move fast without compromising the things that matter. Security: Security-minded by default. You design with the threat model in mind, not as an afterthought. What We Value Intellectual horsepower. Quickly grasping complex technical and business concepts. Kindness and integrity. Earning trust is central to how we build relationships with customers and colleagues. Bias for action. We move quickly to deliver impact and protect our customers against fast-moving threats. Ownership mindset. You care about uptime and stability the way a founder cares about the product. Compensation The base salary range for this full-time position is $120,000 to $190,000, plus equity and benefits. Nametag is a founding member of the Open Imperative, publicly committed to pay equity in the technology industry. We post positions with ranges to encourage people of different backgrounds and experiences to apply. Every offer is benchmarked against market data to ensure fairness and consistency. Final compensation is determined by role, level, and additional factors such as skills, experience, and education. Your recruiter or hiring manager can share more details during the hiring process. Culture & Perks At Nametag, we believe trust starts with how we treat each other. We're a remote-first team that values autonomy, inclusivity, and collaboration, with regular in-person time to stay connected and innovate together. Remote-first: Work from anywhere in the US. Our team spans Seattle, San Francisco, Ann Arbor, Denver, New York City, and beyond. Off-sites: We bring the team together once per quarter for in-person collaboration, often off-site in new places. Flexible schedules: Work in your own time zone; we align key meetings across a shared window. We Offer Competitive salary Meaningful equity ownership Comprehensive health benefits (medical, dental, vision) Flexible paid time off Quarterly team off-sites and travel support New computer hardware and equipment An inclusive environment where your voice has impact and your work drives change
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
08/05/2026
Full time
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
08/05/2026
Full time
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
08/04/2026
Full time
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
08/04/2026
Full time
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
08/04/2026
Full time
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.
08/04/2026
Full time
AI Platform Engineer AI Platform Engineering Full-Time Hybrid Onsite (3 days/week) The Opportunity MassMutual's AI Platform Engineering team is looking for a curious, motivated AI Platform Engineer to launch their engineering career on a high-performing team. This is an entry-level role built for recent graduates and early-career engineers. You will learn directly from experienced platform engineers, contribute to real initiatives from your first weeks, and steadily build the skills to help design, deploy, and operate the systems that power AI across the enterprise. We care less about everything you already know and more about how quickly you learn, how you approach problems, and how much you care about doing good engineering work. The Team This is a unique opportunity to join the team that builds and operates the AI platform powering MassMutual's AI initiatives. The team works at the intersection of cloud infrastructure, AI/ML systems, and developer experience-delivering the foundational capabilities that shape how the entire organization builds and deploys AI. We partner closely with AI engineering, product, and cloud engineering teams across the enterprise, and we invest in growth through a culture of peer learning, candid feedback, mentorship, and shared technical standards. It is a team where early-career engineers are set up to succeed: hard problems are made tractable through clear documentation, thoughtful onboarding, and engineers who genuinely enjoy teaching. The Impact Contribute to platform components-cloud infrastructure, AI serving layers, and developer tooling-under the guidance of senior engineers, growing your understanding of how the pieces fit together. Support the design and implementation of platform features such as the LLM gateway, model serving infrastructure, and integration patterns-writing code, tests, and documentation with regular feedback from your team. Learn the team's engineering standards by participating in design reviews and code reviews, and by pairing with more experienced engineers on real problems. Take ownership of well-scoped tasks within larger initiatives, delivering them to production with support and steadily taking on more scope over time. Help keep the platform healthy by learning reliability practices-monitoring, alerting, SLOs, and incident reviews-and pitching in on operational work. Build familiarity with governance and compliance concepts such as access management, audit logging, and AI usage policies, and why they matter to enterprise customers. Communicate clearly and ask good questions-sharing what you learn, flagging blockers early, and collaborating with teammates and partner teams. Invest in your own growth through mentorship, pairing, and continuous learning, with the goal of ramping toward greater technical independence. The Minimum Qualifications Bachelor's degree in Computer Science, Software Engineering, or a related technical field Foundational understanding of cloud computing and exposure to at least one major cloud provider (AWS, GCP, or Azure) through coursework, labs, or projects as shown by coursework or certification. 2+ years experience in programming proficiency in at least one language such as Python, Go, Java, or a comparable language (this could include coursework, internship experience, bootcamp, etc). The Ideal Qualifications Professional experience preferred: Internships, co-ops, apprenticeships, academic projects, and substantial personal projects all count. Basic familiarity with version control (Git) and a willingness to learn CI/CD, containers (Docker/Kubernetes), and infrastructure-as-code. Curiosity about AI/ML systems and an interest in how models are deployed and served in production. Strong written and verbal communication and a genuine eagerness to learn from feedback. Hands-on exposure to Kubernetes, Docker, or Terraform through coursework, certifications, hackathons, or personal projects. Relevant entry-level certifications are a plus but not required-for example AWS Certified Cloud Practitioner, AWS Solutions Architect - Associate, or CKAD. Any hands-on experience with AI/ML frameworks or LLM APIs, even at a hobby or class-project level. A public portfolio or open-source contributions (e.g., GitHub) that show initiative and a habit of building. Experience collaborating on a team project such as a capstone, hackathon, or group assignment. Comfort with ambiguity and enthusiasm for learning quickly in a fast-moving space. What to Expect as Part of MassMutual and the Team Structured onboarding and a dedicated mentor to support your ramp-up and early growth Regular meetings with the AI Platform Engineering team Focused one-on-one meetings with your manager Networking opportunities including access to Asian, Hispanic/Latinx, African American, women, LGBTQIA+, veteran, and disability-focused Business Resource Groups Access to learning content on Degreed and other informational platforms A company with a strong and stable ethical business, industry-leading pay and benefits, where your ethics and integrity will be valued MassMutual is an equal employment opportunity employer. We welcome all persons to apply. If you need an accommodation to complete the application process, please contact us and share the specifics of the assistance you need. California residents: For detailed information about your rights under the California Consumer Privacy Act (CCPA), please visit our California Consumer Privacy Act Disclosures page.