it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

186 jobs found

Email me jobs like this
Refine Search
Current Search
sr data engineer
Software Engineering Team Lead
ComTec Solutions Rochester, New York
Description: Software Engineering Team Lead Department: Enterprise Systems Group Billable Hours Goal: 60% of worked hours Position Type: Full Time Travel Required: Minimal as needed JOB SUMMARY: As the Technical Lead of the Enterprise Systems Group, you will lead the technical team to help ensure they are following processes to deliver all technical functions to agreed scope and specifications as well as provide ongoing client support. Streamline and provide onboarding activities by shadowing, training, ensuring training teams are completing training and other activities. Making decisions aimed at improving efficiency while at the same time ensuring no disruption of on-time delivery to customers. Directly working with the consulting leads to provide quality solution design. REPORTS TO: Manager, Enterprise Systems Group DIRECT REPORTS: Systems and Software Engineers ESSENTIAL FUNCTIONS: Utilize Software Engineering knowledge providing customized solutions to Epicor and external utilities Develop and test solutions with attention to detail and accuracy Document all modifications to client software according to company policy Manage assigned queue, to ensure the team is meeting deadlines and other milestones Provide classroom training to end users on Epicor technical tools Mentor Technical Team Members as required Assist in the hiring process of Technical Team Members Assist with onboarding/training of new Technical Team Members Actively contribute to development team process improvement, policies and procedures Provide direction and escalation for ticket / task issues, and customer satisfaction follow up Present ComTec in a positive light both internally and externally (i.e., be a champion for the company) Assist with performance and development of Technical Team Members Develop and monitor yearly training plans for direct reports Assist Director with quarterly performance reviews with direct reports ADDITIONAL RESPONSIBILITIES: Troubleshoot, identify and evaluate alternative solutions to a problem Maintain daily timesheet and expense report entries and submit them accurately and timely Lead weekly team meetings. Attend Epicor Escalation Meetings Other duties as required Requirements: TECHNICAL SKILLS: C# / VB.NET (intermediate/Advanced) Knowledge of Microsoft SQL Server and/or Progress Databases (Intermediate) Crystal Reports development (Basic) Microsoft SSRS Reporting (Basic) SOFT SKILLS & ABILITIES: Must be able to read, correctly interpret, develop, implement and test solutions based on the specifications document Strong written and verbal communication skills Pleasant and professional demeanor in all client and internal communications Intellectually resourceful with sound judgment and effective decision-making abilities Independent worker and able to work effectively on daily tasks without direct supervision Strong organization skills and ability to operate efficiently throughout daily tasks In general, owns issues through resolution although understands when to escalate a problem to another team member and whom to escalate too; accepts escalated issues; and mentors when appropriate Demonstrate empathy and professionalism with users at all times Work well with clients at all levels Operate with client satisfaction in mind Energy, enthusiasm and results-oriented EDUCATION, EXPERIENCE, & KNOWLEDGE: Related bachelor's degree or equivalent work experience 4+ years of Microsoft .NET programming experience Epicor application experience a plus WORK ENVIRONMENT/PHYSICAL DEMANDS: Use of computer and office equipment Ability to remain calm in stressful situations Performs all administrative functions expected at this level ADDITIONAL REQUIREMENTS: Ability to schedule for evening or weekend work occasionally Valid driver's license in your state of residence and reliable personal vehicle Compensation details: 00 Yearly Salary PIc08683a6d6-
08/05/2026
Full time
Description: Software Engineering Team Lead Department: Enterprise Systems Group Billable Hours Goal: 60% of worked hours Position Type: Full Time Travel Required: Minimal as needed JOB SUMMARY: As the Technical Lead of the Enterprise Systems Group, you will lead the technical team to help ensure they are following processes to deliver all technical functions to agreed scope and specifications as well as provide ongoing client support. Streamline and provide onboarding activities by shadowing, training, ensuring training teams are completing training and other activities. Making decisions aimed at improving efficiency while at the same time ensuring no disruption of on-time delivery to customers. Directly working with the consulting leads to provide quality solution design. REPORTS TO: Manager, Enterprise Systems Group DIRECT REPORTS: Systems and Software Engineers ESSENTIAL FUNCTIONS: Utilize Software Engineering knowledge providing customized solutions to Epicor and external utilities Develop and test solutions with attention to detail and accuracy Document all modifications to client software according to company policy Manage assigned queue, to ensure the team is meeting deadlines and other milestones Provide classroom training to end users on Epicor technical tools Mentor Technical Team Members as required Assist in the hiring process of Technical Team Members Assist with onboarding/training of new Technical Team Members Actively contribute to development team process improvement, policies and procedures Provide direction and escalation for ticket / task issues, and customer satisfaction follow up Present ComTec in a positive light both internally and externally (i.e., be a champion for the company) Assist with performance and development of Technical Team Members Develop and monitor yearly training plans for direct reports Assist Director with quarterly performance reviews with direct reports ADDITIONAL RESPONSIBILITIES: Troubleshoot, identify and evaluate alternative solutions to a problem Maintain daily timesheet and expense report entries and submit them accurately and timely Lead weekly team meetings. Attend Epicor Escalation Meetings Other duties as required Requirements: TECHNICAL SKILLS: C# / VB.NET (intermediate/Advanced) Knowledge of Microsoft SQL Server and/or Progress Databases (Intermediate) Crystal Reports development (Basic) Microsoft SSRS Reporting (Basic) SOFT SKILLS & ABILITIES: Must be able to read, correctly interpret, develop, implement and test solutions based on the specifications document Strong written and verbal communication skills Pleasant and professional demeanor in all client and internal communications Intellectually resourceful with sound judgment and effective decision-making abilities Independent worker and able to work effectively on daily tasks without direct supervision Strong organization skills and ability to operate efficiently throughout daily tasks In general, owns issues through resolution although understands when to escalate a problem to another team member and whom to escalate too; accepts escalated issues; and mentors when appropriate Demonstrate empathy and professionalism with users at all times Work well with clients at all levels Operate with client satisfaction in mind Energy, enthusiasm and results-oriented EDUCATION, EXPERIENCE, & KNOWLEDGE: Related bachelor's degree or equivalent work experience 4+ years of Microsoft .NET programming experience Epicor application experience a plus WORK ENVIRONMENT/PHYSICAL DEMANDS: Use of computer and office equipment Ability to remain calm in stressful situations Performs all administrative functions expected at this level ADDITIONAL REQUIREMENTS: Ability to schedule for evening or weekend work occasionally Valid driver's license in your state of residence and reliable personal vehicle Compensation details: 00 Yearly Salary PIc08683a6d6-
Senior Cloud Engineer
Gridware San Francisco, California
Job Description Job Description About Gridware Gridware is a San Francisco-based technology company dedicated to protecting and enhancing the electrical grid. We pioneered a groundbreaking new class of grid management called active grid response (AGR), focused on monitoring the electrical, physical, and environmental aspects of the grid that affect reliability and safety. Gridware's advanced Active Grid Response platform uses high-precision sensors to detect potential issues early, enabling proactive maintenance and fault mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the grid operates efficiently. The company is backed by climate-tech and Silicon Valley investors. For more information, please visit . Role Description We're scaling the deployment of critical infrastructure monitoring devices to detect real-world fault events that lead to wildfires. The platform you'll build and operate ingests millions of events per day from devices in the field, powers customer-facing dashboards and alerting, and supports the data science work that turns raw signals into grid intelligence. You will own AWS infrastructure, Kubernetes (EKS), CI/CD, and observability end-to-end, partnering with our Cloud Security team to keep the platform safe and compliant, and with backend, firmware, and data teams to keep them shipping fast. As an early member of the DevOps team, you'll have a direct hand in shaping how Gridware builds, deploys, and runs production systems for years to come. Responsibilities Design, build, and operate scalable, secure, and highly available cloud infrastructure across AWS. Own and evolve our Kubernetes platform, enabling reliable application deployment and operations through GitOps best practices. Build and maintain CI/CD systems that improve developer velocity, release quality, and operational reliability. Manage and optimize event-driven infrastructure powering high-volume telemetry and device data pipelines. Define and maintain Infrastructure as Code standards, ensuring consistency, repeatability, and scalability across environments. Develop and enhance observability, monitoring, and incident response capabilities to support reliable production operations. Partner closely with Security and Engineering teams to strengthen platform security, access management, and operational resilience. Troubleshoot complex production issues, drive root cause analysis, and turn lessons learned into automation, tooling, and operational improvements. Required Skills 5+ years of experience in DevOps, SRE, or Platform Engineering operating production AWS environments Deep expertise with Kubernetes (EKS preferred), GitOps workflows (Argo CD/Flux), and Infrastructure as Code (Terraform) Strong experience building and maintaining CI/CD pipelines, ideally with GitHub Actions Hands-on experience operating distributed systems and cloud-native platforms (e.g., Kafka/MSK) Solid understanding of networking, DNS, TLS, identity/access management, and cloud security best practices Experience with observability, monitoring, and logging tools such as Grafana, Prometheus, Loki, or similar Strong Linux, scripting, and troubleshooting skills with the ability to debug complex production issues end-to-end Bonus Skills Experience operating Apollo Router / GraphQL federation gateways in production. Experience operating Argo Workflows or similar Kubernetes-native job / pipeline runners in production. Familiarity with Databricks or ML Ops pipelines for data and model deployment. Experience designing, operating, and exercising Disaster Recovery (DR) environments, including cross-region replication, backups, and tested failover runbooks. Experience with Tailscale or other zero-trust networking tools. Experience supporting IoT / embedded fleets at scale, including secure device-to-cloud connectivity. Experience in high-growth startup environments where you must wear many hats. This describes the ideal candidate; many of us have picked up this expertise along the way. Even if you meet only part of this list, we encourage you to apply! Gridware Technologies Inc. is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to any characteristic protected by applicable federal, state, or local law. Benefits Health, Dental & Vision (Gold and Platinum with some providers plans fully covered) Paid parental leave Alternating day off (every other Monday) "Off the Grid", a two week per year paid break for all employees. Commuter allowance Company-paid training
08/05/2026
Full time
Job Description Job Description About Gridware Gridware is a San Francisco-based technology company dedicated to protecting and enhancing the electrical grid. We pioneered a groundbreaking new class of grid management called active grid response (AGR), focused on monitoring the electrical, physical, and environmental aspects of the grid that affect reliability and safety. Gridware's advanced Active Grid Response platform uses high-precision sensors to detect potential issues early, enabling proactive maintenance and fault mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the grid operates efficiently. The company is backed by climate-tech and Silicon Valley investors. For more information, please visit . Role Description We're scaling the deployment of critical infrastructure monitoring devices to detect real-world fault events that lead to wildfires. The platform you'll build and operate ingests millions of events per day from devices in the field, powers customer-facing dashboards and alerting, and supports the data science work that turns raw signals into grid intelligence. You will own AWS infrastructure, Kubernetes (EKS), CI/CD, and observability end-to-end, partnering with our Cloud Security team to keep the platform safe and compliant, and with backend, firmware, and data teams to keep them shipping fast. As an early member of the DevOps team, you'll have a direct hand in shaping how Gridware builds, deploys, and runs production systems for years to come. Responsibilities Design, build, and operate scalable, secure, and highly available cloud infrastructure across AWS. Own and evolve our Kubernetes platform, enabling reliable application deployment and operations through GitOps best practices. Build and maintain CI/CD systems that improve developer velocity, release quality, and operational reliability. Manage and optimize event-driven infrastructure powering high-volume telemetry and device data pipelines. Define and maintain Infrastructure as Code standards, ensuring consistency, repeatability, and scalability across environments. Develop and enhance observability, monitoring, and incident response capabilities to support reliable production operations. Partner closely with Security and Engineering teams to strengthen platform security, access management, and operational resilience. Troubleshoot complex production issues, drive root cause analysis, and turn lessons learned into automation, tooling, and operational improvements. Required Skills 5+ years of experience in DevOps, SRE, or Platform Engineering operating production AWS environments Deep expertise with Kubernetes (EKS preferred), GitOps workflows (Argo CD/Flux), and Infrastructure as Code (Terraform) Strong experience building and maintaining CI/CD pipelines, ideally with GitHub Actions Hands-on experience operating distributed systems and cloud-native platforms (e.g., Kafka/MSK) Solid understanding of networking, DNS, TLS, identity/access management, and cloud security best practices Experience with observability, monitoring, and logging tools such as Grafana, Prometheus, Loki, or similar Strong Linux, scripting, and troubleshooting skills with the ability to debug complex production issues end-to-end Bonus Skills Experience operating Apollo Router / GraphQL federation gateways in production. Experience operating Argo Workflows or similar Kubernetes-native job / pipeline runners in production. Familiarity with Databricks or ML Ops pipelines for data and model deployment. Experience designing, operating, and exercising Disaster Recovery (DR) environments, including cross-region replication, backups, and tested failover runbooks. Experience with Tailscale or other zero-trust networking tools. Experience supporting IoT / embedded fleets at scale, including secure device-to-cloud connectivity. Experience in high-growth startup environments where you must wear many hats. This describes the ideal candidate; many of us have picked up this expertise along the way. Even if you meet only part of this list, we encourage you to apply! Gridware Technologies Inc. is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to any characteristic protected by applicable federal, state, or local law. Benefits Health, Dental & Vision (Gold and Platinum with some providers plans fully covered) Paid parental leave Alternating day off (every other Monday) "Off the Grid", a two week per year paid break for all employees. Commuter allowance Company-paid training
Part-Time Student Worker - Hardware Software Integration Engineer
Zoox Hayward, California
Job Description Job Description About Zoox Zoox is an autonomous ride-hailing company building the world's first purpose-built robotaxi - fully electric, bidirectional, with no steering wheel or driver's seat. Backed by Amazon and founded to make transportation safer, cleaner, and more accessible, Zoox designs its vehicles entirely around the rider. We're currently operating in Las Vegas and San Francisco, with Austin and Miami on the horizon, and testing underway across seven U.S. markets. About Our Part-Time Student Worker Program Zoox's part-time student worker program puts you at the center of one of the most ambitious challenges in transportation. You'll contribute to real projects, work alongside engineers and researchers pushing the boundaries of autonomous technology, and gain experience that goes well beyond the classroom. We're looking for students who bring strong academic foundations, curiosity that doesn't stop at coursework, and a drive to be part of something that matters. Responsibilities Execute and coordinate validation cycles for Zoox's Manufacturing test infrastructure with a focus on driving test infra stability Act as on-call support for the factory, representing the HWSI-FW team as you document, triage, characterize and report stability issues in order to unblock factory operations Program Requirements Currently enrolled in a B.S. or M.S. in Automotive Engineering, Robotics, Systems Engineering, Mechatronics, or a relevant technical field Available to commit to a minimum three-month assignment Able to commit to a minimum of 20 hours per week Able to work on-site at one of our office locations Must adhere with Zoox confidentiality requirements, including refraining from using or sharing proprietary company information outside of Zoox, such as in academic research, theses, publications, or presentations Qualifications Ability to understand and navigate complex technical systems Ability to grapple with ambiguity and collaborate with cross-functional teams Experience with Linux, Git, and Jira Bonus Qualifications Experience with autonomous vehicles, robotics, or other safety-critical systems Experience with simulation environments and hardware-in-the-loop (HIL) testing We want to be transparent: this is not an internship. The Part-Time Student Worker Program is designed to complement your academic experience by providing meaningful, ongoing work alongside your studies. Rather than participating in a cohort-based program, you'll join a team directly and contribute to real projects with real impact. While the program does not include structured intern programming or a pathway to full-time employment, it offers valuable opportunities to learn, develop new skills, and gain hands-on experience in a professional environment. Compensation for this role is $30/hour. This is a contract position; employment will be through a vendor contracted with Zoox. The hourly rate is as posted, and benefits eligibility is determined by the vendor. About Zoox Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We're looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team. Follow us on LinkedIn Accommodations If you need an accommodation to participate in the application or interview process please reach out to or your assigned recruiter. A Final Note: You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
08/05/2026
Full time
Job Description Job Description About Zoox Zoox is an autonomous ride-hailing company building the world's first purpose-built robotaxi - fully electric, bidirectional, with no steering wheel or driver's seat. Backed by Amazon and founded to make transportation safer, cleaner, and more accessible, Zoox designs its vehicles entirely around the rider. We're currently operating in Las Vegas and San Francisco, with Austin and Miami on the horizon, and testing underway across seven U.S. markets. About Our Part-Time Student Worker Program Zoox's part-time student worker program puts you at the center of one of the most ambitious challenges in transportation. You'll contribute to real projects, work alongside engineers and researchers pushing the boundaries of autonomous technology, and gain experience that goes well beyond the classroom. We're looking for students who bring strong academic foundations, curiosity that doesn't stop at coursework, and a drive to be part of something that matters. Responsibilities Execute and coordinate validation cycles for Zoox's Manufacturing test infrastructure with a focus on driving test infra stability Act as on-call support for the factory, representing the HWSI-FW team as you document, triage, characterize and report stability issues in order to unblock factory operations Program Requirements Currently enrolled in a B.S. or M.S. in Automotive Engineering, Robotics, Systems Engineering, Mechatronics, or a relevant technical field Available to commit to a minimum three-month assignment Able to commit to a minimum of 20 hours per week Able to work on-site at one of our office locations Must adhere with Zoox confidentiality requirements, including refraining from using or sharing proprietary company information outside of Zoox, such as in academic research, theses, publications, or presentations Qualifications Ability to understand and navigate complex technical systems Ability to grapple with ambiguity and collaborate with cross-functional teams Experience with Linux, Git, and Jira Bonus Qualifications Experience with autonomous vehicles, robotics, or other safety-critical systems Experience with simulation environments and hardware-in-the-loop (HIL) testing We want to be transparent: this is not an internship. The Part-Time Student Worker Program is designed to complement your academic experience by providing meaningful, ongoing work alongside your studies. Rather than participating in a cohort-based program, you'll join a team directly and contribute to real projects with real impact. While the program does not include structured intern programming or a pathway to full-time employment, it offers valuable opportunities to learn, develop new skills, and gain hands-on experience in a professional environment. Compensation for this role is $30/hour. This is a contract position; employment will be through a vendor contracted with Zoox. The hourly rate is as posted, and benefits eligibility is determined by the vendor. About Zoox Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We're looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team. Follow us on LinkedIn Accommodations If you need an accommodation to participate in the application or interview process please reach out to or your assigned recruiter. A Final Note: You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Sr. Machine Learning Engineer (Perception and Tracking)
Ouster San Francisco, California
Job Description Job Description At Ouster, we build sensors and tools for engineers, roboticists, and researchers, so they can make the world safer and more efficient. We've transformed LIDAR from an analog device with thousands of components to an elegant digital device powered by one chip-scale laser array and one CMOS sensor. The result is a full range of high-resolution LIDAR sensors that deliver superior imaging at a dramatically lower price. Our advanced sensor hardware and vision algorithms are used in autonomous cars, robotics, industrial, and smart infrastructure applications (among many others). If you're motivated by solving big problems, we're hiring key roles across the company and need your help! We are looking for a highly technical Machine Learning Engineer to lead our efforts in Object Detection and Tracking. You will not simply be "importing" pre-made models; you will be architecting deep neural networks, translating state-of-the-art research papers into code, and optimizing these systems for real-time, on-device performance. This role requires a deep knowledge of neural network architectures. You should be confident ripping apart a model to modify layers, loss functions, and data flows to fit our specific constraints. Key Responsibilities Architect Unified Models: Design and train DNN models that perform Object Detection and Tracking simultaneously, leveraging temporal information to improve consistency. Research to Production: Evaluate state-of-the-art research papers and prototype these concepts (turning papers into code) and adapt them into robust, production-grade solutions. Deep Model Customization: Go beyond standard libraries by implementing custom loss functions, modifying internal model architectures, and designing specific data augmentation strategies to squeeze out maximum performance. Edge Optimization: Ensure high accuracy is matched by high efficiency. Optimize models for real-time inference and on-device deployment. Data Strategy: Develop training recipes for data-constrained environments and effective post-training strategies. Required Qualifications Core Stack: 5+ years proficiency in Python and PyTorch. 3+ years proficiency in C++ for production deployment and optimization. Detection & Tracking: Deep theoretical and practical understanding of modern object detectors (e.g., Transformers, YOLO variants, R-CNNs) and tracking algorithms (e.g., DeepSORT, Kalman Filters, Optical Flow). Architecture Internals: Proven experience not being dependent on "out-of-the-box" APIs. You have a track record of modifying model architectures via extensive experimentation to meet specific requirements. Low-Data Regimes: Experience improving model generalization with limited data using Transfer Learning, Domain Adaptation, or Few-Shot Learning. Mathematical Foundation: Strong grasp of linear algebra and probability as it applies to custom loss function design and geometric 3D vision. Preferred Qualifications 3D / LiDAR Experience: Hands-on experience with 3D Point Cloud data (LiDAR) is a massive plus. Deployment Tools: Experience with TensorRT, ONNX Runtime, or edge-specific hardware (NVIDIA Jetson, etc.). The base pay will be dependent on your skills, work experience, location, and qualifications. This role may also be eligible for equity & benefits. ($180,000-220,000) We acknowledge the confidence gap at Ouster. You do not need to meet all of these requirements to be the ideal candidate for this role. Ouster is an Equal Employment Opportunity employer that pursues and hires a diverse workforce. Ouster does not make employment decisions on the basis of race, color, religion, ethnic or national origin, nationality, sex, gender, gender-identity, sexual orientation, disability, age, military status, or any other basis protected by local, state, or federal laws. Ouster also strives for a healthy and safe workplace, and prohibits harassment of any kind. Pursuant to the San Francisco Fair Chance Ordinance, Ouster considers qualified applicants with arrest and conviction records for employment. If you have a disability or special need that requires accommodation, please let us know. Powered by JazzHR POaf4TokKv
08/05/2026
Full time
Job Description Job Description At Ouster, we build sensors and tools for engineers, roboticists, and researchers, so they can make the world safer and more efficient. We've transformed LIDAR from an analog device with thousands of components to an elegant digital device powered by one chip-scale laser array and one CMOS sensor. The result is a full range of high-resolution LIDAR sensors that deliver superior imaging at a dramatically lower price. Our advanced sensor hardware and vision algorithms are used in autonomous cars, robotics, industrial, and smart infrastructure applications (among many others). If you're motivated by solving big problems, we're hiring key roles across the company and need your help! We are looking for a highly technical Machine Learning Engineer to lead our efforts in Object Detection and Tracking. You will not simply be "importing" pre-made models; you will be architecting deep neural networks, translating state-of-the-art research papers into code, and optimizing these systems for real-time, on-device performance. This role requires a deep knowledge of neural network architectures. You should be confident ripping apart a model to modify layers, loss functions, and data flows to fit our specific constraints. Key Responsibilities Architect Unified Models: Design and train DNN models that perform Object Detection and Tracking simultaneously, leveraging temporal information to improve consistency. Research to Production: Evaluate state-of-the-art research papers and prototype these concepts (turning papers into code) and adapt them into robust, production-grade solutions. Deep Model Customization: Go beyond standard libraries by implementing custom loss functions, modifying internal model architectures, and designing specific data augmentation strategies to squeeze out maximum performance. Edge Optimization: Ensure high accuracy is matched by high efficiency. Optimize models for real-time inference and on-device deployment. Data Strategy: Develop training recipes for data-constrained environments and effective post-training strategies. Required Qualifications Core Stack: 5+ years proficiency in Python and PyTorch. 3+ years proficiency in C++ for production deployment and optimization. Detection & Tracking: Deep theoretical and practical understanding of modern object detectors (e.g., Transformers, YOLO variants, R-CNNs) and tracking algorithms (e.g., DeepSORT, Kalman Filters, Optical Flow). Architecture Internals: Proven experience not being dependent on "out-of-the-box" APIs. You have a track record of modifying model architectures via extensive experimentation to meet specific requirements. Low-Data Regimes: Experience improving model generalization with limited data using Transfer Learning, Domain Adaptation, or Few-Shot Learning. Mathematical Foundation: Strong grasp of linear algebra and probability as it applies to custom loss function design and geometric 3D vision. Preferred Qualifications 3D / LiDAR Experience: Hands-on experience with 3D Point Cloud data (LiDAR) is a massive plus. Deployment Tools: Experience with TensorRT, ONNX Runtime, or edge-specific hardware (NVIDIA Jetson, etc.). The base pay will be dependent on your skills, work experience, location, and qualifications. This role may also be eligible for equity & benefits. ($180,000-220,000) We acknowledge the confidence gap at Ouster. You do not need to meet all of these requirements to be the ideal candidate for this role. Ouster is an Equal Employment Opportunity employer that pursues and hires a diverse workforce. Ouster does not make employment decisions on the basis of race, color, religion, ethnic or national origin, nationality, sex, gender, gender-identity, sexual orientation, disability, age, military status, or any other basis protected by local, state, or federal laws. Ouster also strives for a healthy and safe workplace, and prohibits harassment of any kind. Pursuant to the San Francisco Fair Chance Ordinance, Ouster considers qualified applicants with arrest and conviction records for employment. If you have a disability or special need that requires accommodation, please let us know. Powered by JazzHR POaf4TokKv
Sr. Finance Reporting Engineer
Align Technology San Jose, California
Job Description Job Description Description This position is ideal for senior-level professionals to join the Finance Systems & Reporting team as a Sr. Finance Reporting Engineer. The Finance Engineer will be reporting to the Director, Global Finance Systems & Reporting based in San Jose, California (US) and would strategically engage with key members of finance function defining and delivering on the enterprise roadmap projects as well as own and help resolve day to day tactical problems and challenges. The Sr. Finance Reporting Engineer will be responsible for designing, developing, and maintaining enterprise financial reporting and analytics solutions. This role focuses on delivering accurate, scalable, and automated reporting capabilities by leveraging SAP technologies, cloud-based data platforms, and business intelligence tools to support strategic decision-making across the organization. Role expectations Lead the financial reporting life cycle end-to-end that includes conducting requirement gathering sessions with business, design the reporting solution, development, deployment, maintenance and training of the reporting solution to broader regional audience. Oversee the design and development for the planning system, OneStream, and be prepared to use other planning and reporting tools used at Align, including ERP, Consolidation Planning, and Business Intelligence software packages. Support financial close activities through data reconciliation, validation, and coordination of data loads for month-end, quarter-end, and year-end reporting. Responsible for designing, developing, and maintaining scalable data pipelines and integration workflows to ensure accurate and efficient data processing across systems. Monitor data flows, task chains, and job performance; implementing data quality controls and validation processes; troubleshooting and resolving data load issues; and optimizing data transformation and loading performance. Develop and troubleshoot a variety of Onestream components that include transformation rules, security models, custom dashboards, reports, business rules, and member formulas, with a focus on enhancing system functionality and efficiency. Lead the collaboration and development efforts for various financial Power BI dashboards required for finance, that will make analytics easier and reporting simplified for business at the same time maintaining data accuracy standards as required for financial reporting. Acting as a liaison between business and IT groups. Work with IT to help them understand business requirements and translate it in technical terminologies for them. Help business understand the technical solutions deployed and train them on how they can effectively use it for their needs. Identifying the gaps with existing solutions in place, finding out solutions to resolve them and work with IT/business to get them implemented. Lead the integration testing and user acceptance testing with business and IT collaboratively. Key Personnel responsible for financial close reporting of critical reports Key person responsible for ensuring financial data integrity and completeness for all finance data that will be needed for financial close reporting. Comply with data security and access control standards, maintains thorough process documentation, and provides production support, including participation in on-call rotations with the offshore team when required. Assist with special projects and ad-hoc requests, as necessary. What we're looking for Requires Bachelor's degree in CS, Information Technology, or a related field 12+ years of experience in Data loading and Monitoring systems to support the Finance organization. Strong understanding of financial reporting, FP&A, close processes, and management reporting. Hands-on experience with SAP Datasphere for data modeling, data warehousing, integration, and governance across SAP and non-SAP landscapes. Strong understanding of SAP RTR business processes, hands on experience working with FI-GL, COPA, AR/AP, Should be a self-starter who is able to work with minimal direction and exercises considerable latitude in determining objectives and approaches to assignments. Hands-on experience working with Onestream planning system or similar planning platforms such as Anaplan, Planful. Hands-on experience with SQL and object-oriented (VB.Net, C#) coding experience is preferred. The candidate will serve as a liaison between the finance user group, corporate report development team, and IT. Should be a team player and possess good interpersonal and communication skills, reflecting an ability to be patient and outgoing with people. Should be highly motivated, result focused, and act with a high sense of urgency. Should possess excellent planning and prioritization skills with the ability to multitask and maintain by adapting to change. Complementary skills Advanced Microsoft Outlook, Word, Excel and PowerPoint skills. Must have the ability to independently create spreadsheets and perform quantitative analysis. Prior experience working with Azure Datalake, S/4 HANA, SAP-ECC or SAP BW is a plus. Pay Transparency If provided, base salary or wage rate ranges are the range in which Align reasonably expects to set a candidate's pay for the posted position. Actual placement depends on the individual skills and experience level of a candidate plus the total compensation and equity across team members. For other locations outside of the primary location, the base salary range will be adjusted geographically. For Field Sales roles, the salary listed is the base pay only and does not include the applicable incentive compensation plan. A cost of living adjustment may be added to base pay for higher cost areas in the U.S. Our internship hourly rates are a standard pay determined based on the position and your location, year in school, degree, and experience. General Description of All Benefits We are pleased to provide a general description of the benefits Align offers to full-time employees in this position. Family Benefits. Align offers employees and their eligible dependents medical (with a Health Savings Account option for some plan offerings), dental, and vision in accordance with those plans. Align also offers to employees: Discounts on Invisalign and Vivera to employees and their eligible dependents after 90 days of employment Back-up Child/Elder Care and access to a caregiving concierge Family Forming Benefits - Available to Employees, and their spouse or domestic partner, covered under one of Align's health plans Breast Milk Delivery and Lactation Support Services Employee Assistance Program Hinge Health Virtual Physical Therapy - Available to all employees and eligible dependents (age 18+) enrolled in an Align medical Plan Employee benefits. Align offers its employees: Short-term and long-term disability insurance in accordance with those plans. Basic Life Insurance and Accidental Death and Dismemberment. Voluntary Supplemental Life Insurance for Employee, Spouse/Domestic Partner, and Child(ren) are available for purchase in accordance with those plans. Flexible Spending Accounts- Employees may be eligible to participate in a health care account (including a limited health FSA if enrolled in a HDHP), dependent care account, and a pre-tax commuter benefit plan. 401k plan (with a discretionary Company match of 50% up to 6% of eligible earnings up to a maximum match of 3%.). Employer match vests after two years - 25% year one and 100% at year two. Align offers traditional, Roth, and after-tax options. Employee Stock Purchase Program (Employees must work 20 hours or more and be employed on purchase date to be eligible). Paid vacation of up to 17 days during the first full year of employment (currently accrued at the rate of 5.24 hours each pay-period), which carries over to a maximum cap of 30 days. Annual paid vacation time accrual increases based on tenure. Both exempt and non-exempt employees who work 32 hours or more per week receive prorated vacation accrual based on their regularly scheduled work hours and tenure. Sick time is accrued throughout the year at the rate of one hour for every thirty worked. Employees can carry over unused sick leave each year, up to a maximum balance of 80 hours. 11 Company-designated paid holidays throughout the year. If employed for at least 12 consecutive months, Align will grant up to 6 weeks of paid Parental Leave. If employed for less than 12 consecutive months, Align will grant up to 4 weeks of paid Parental Leave. All parental leave must be completed within one year of the birth or placement of the child. Parental leave is in addition to any state and/or local parental leave benefits. Three days of paid bereavement leave. In some cases, due to travel the amount of paid leave may be extended to 5 paid days off. To the extent applicable state or local law offers more generous benefits, Align complies with any such law. Non-exempt employees will receive full pay for up to 10 days of jury duty. Exempt employees will receive their full salary during any week they serve and perform any work. Other insurance such as legal, critical illness, voluntary accident, long-term care, auto, home and pet insurance are available for purchase. To the extent applicable state or local law offers more generous benefits, Align complies with any such law.
08/05/2026
Full time
Job Description Job Description Description This position is ideal for senior-level professionals to join the Finance Systems & Reporting team as a Sr. Finance Reporting Engineer. The Finance Engineer will be reporting to the Director, Global Finance Systems & Reporting based in San Jose, California (US) and would strategically engage with key members of finance function defining and delivering on the enterprise roadmap projects as well as own and help resolve day to day tactical problems and challenges. The Sr. Finance Reporting Engineer will be responsible for designing, developing, and maintaining enterprise financial reporting and analytics solutions. This role focuses on delivering accurate, scalable, and automated reporting capabilities by leveraging SAP technologies, cloud-based data platforms, and business intelligence tools to support strategic decision-making across the organization. Role expectations Lead the financial reporting life cycle end-to-end that includes conducting requirement gathering sessions with business, design the reporting solution, development, deployment, maintenance and training of the reporting solution to broader regional audience. Oversee the design and development for the planning system, OneStream, and be prepared to use other planning and reporting tools used at Align, including ERP, Consolidation Planning, and Business Intelligence software packages. Support financial close activities through data reconciliation, validation, and coordination of data loads for month-end, quarter-end, and year-end reporting. Responsible for designing, developing, and maintaining scalable data pipelines and integration workflows to ensure accurate and efficient data processing across systems. Monitor data flows, task chains, and job performance; implementing data quality controls and validation processes; troubleshooting and resolving data load issues; and optimizing data transformation and loading performance. Develop and troubleshoot a variety of Onestream components that include transformation rules, security models, custom dashboards, reports, business rules, and member formulas, with a focus on enhancing system functionality and efficiency. Lead the collaboration and development efforts for various financial Power BI dashboards required for finance, that will make analytics easier and reporting simplified for business at the same time maintaining data accuracy standards as required for financial reporting. Acting as a liaison between business and IT groups. Work with IT to help them understand business requirements and translate it in technical terminologies for them. Help business understand the technical solutions deployed and train them on how they can effectively use it for their needs. Identifying the gaps with existing solutions in place, finding out solutions to resolve them and work with IT/business to get them implemented. Lead the integration testing and user acceptance testing with business and IT collaboratively. Key Personnel responsible for financial close reporting of critical reports Key person responsible for ensuring financial data integrity and completeness for all finance data that will be needed for financial close reporting. Comply with data security and access control standards, maintains thorough process documentation, and provides production support, including participation in on-call rotations with the offshore team when required. Assist with special projects and ad-hoc requests, as necessary. What we're looking for Requires Bachelor's degree in CS, Information Technology, or a related field 12+ years of experience in Data loading and Monitoring systems to support the Finance organization. Strong understanding of financial reporting, FP&A, close processes, and management reporting. Hands-on experience with SAP Datasphere for data modeling, data warehousing, integration, and governance across SAP and non-SAP landscapes. Strong understanding of SAP RTR business processes, hands on experience working with FI-GL, COPA, AR/AP, Should be a self-starter who is able to work with minimal direction and exercises considerable latitude in determining objectives and approaches to assignments. Hands-on experience working with Onestream planning system or similar planning platforms such as Anaplan, Planful. Hands-on experience with SQL and object-oriented (VB.Net, C#) coding experience is preferred. The candidate will serve as a liaison between the finance user group, corporate report development team, and IT. Should be a team player and possess good interpersonal and communication skills, reflecting an ability to be patient and outgoing with people. Should be highly motivated, result focused, and act with a high sense of urgency. Should possess excellent planning and prioritization skills with the ability to multitask and maintain by adapting to change. Complementary skills Advanced Microsoft Outlook, Word, Excel and PowerPoint skills. Must have the ability to independently create spreadsheets and perform quantitative analysis. Prior experience working with Azure Datalake, S/4 HANA, SAP-ECC or SAP BW is a plus. Pay Transparency If provided, base salary or wage rate ranges are the range in which Align reasonably expects to set a candidate's pay for the posted position. Actual placement depends on the individual skills and experience level of a candidate plus the total compensation and equity across team members. For other locations outside of the primary location, the base salary range will be adjusted geographically. For Field Sales roles, the salary listed is the base pay only and does not include the applicable incentive compensation plan. A cost of living adjustment may be added to base pay for higher cost areas in the U.S. Our internship hourly rates are a standard pay determined based on the position and your location, year in school, degree, and experience. General Description of All Benefits We are pleased to provide a general description of the benefits Align offers to full-time employees in this position. Family Benefits. Align offers employees and their eligible dependents medical (with a Health Savings Account option for some plan offerings), dental, and vision in accordance with those plans. Align also offers to employees: Discounts on Invisalign and Vivera to employees and their eligible dependents after 90 days of employment Back-up Child/Elder Care and access to a caregiving concierge Family Forming Benefits - Available to Employees, and their spouse or domestic partner, covered under one of Align's health plans Breast Milk Delivery and Lactation Support Services Employee Assistance Program Hinge Health Virtual Physical Therapy - Available to all employees and eligible dependents (age 18+) enrolled in an Align medical Plan Employee benefits. Align offers its employees: Short-term and long-term disability insurance in accordance with those plans. Basic Life Insurance and Accidental Death and Dismemberment. Voluntary Supplemental Life Insurance for Employee, Spouse/Domestic Partner, and Child(ren) are available for purchase in accordance with those plans. Flexible Spending Accounts- Employees may be eligible to participate in a health care account (including a limited health FSA if enrolled in a HDHP), dependent care account, and a pre-tax commuter benefit plan. 401k plan (with a discretionary Company match of 50% up to 6% of eligible earnings up to a maximum match of 3%.). Employer match vests after two years - 25% year one and 100% at year two. Align offers traditional, Roth, and after-tax options. Employee Stock Purchase Program (Employees must work 20 hours or more and be employed on purchase date to be eligible). Paid vacation of up to 17 days during the first full year of employment (currently accrued at the rate of 5.24 hours each pay-period), which carries over to a maximum cap of 30 days. Annual paid vacation time accrual increases based on tenure. Both exempt and non-exempt employees who work 32 hours or more per week receive prorated vacation accrual based on their regularly scheduled work hours and tenure. Sick time is accrued throughout the year at the rate of one hour for every thirty worked. Employees can carry over unused sick leave each year, up to a maximum balance of 80 hours. 11 Company-designated paid holidays throughout the year. If employed for at least 12 consecutive months, Align will grant up to 6 weeks of paid Parental Leave. If employed for less than 12 consecutive months, Align will grant up to 4 weeks of paid Parental Leave. All parental leave must be completed within one year of the birth or placement of the child. Parental leave is in addition to any state and/or local parental leave benefits. Three days of paid bereavement leave. In some cases, due to travel the amount of paid leave may be extended to 5 paid days off. To the extent applicable state or local law offers more generous benefits, Align complies with any such law. Non-exempt employees will receive full pay for up to 10 days of jury duty. Exempt employees will receive their full salary during any week they serve and perform any work. Other insurance such as legal, critical illness, voluntary accident, long-term care, auto, home and pet insurance are available for purchase. To the extent applicable state or local law offers more generous benefits, Align complies with any such law.
Senior Site Reliability Engineer
ASAPP Mountain View, California
Job Description Job Description At ASAPP, our mission is simple: deliver the best AI-powered customer experience-faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed, ownership, and a relentless focus on outcomes. We work in tight, skilled teams, prioritize clarity over complexity, and continuously evolve through curiosity, data, and craftsmanship. We're seeking technologists and problem solvers who thrive in fast-paced environments, love collaborating with great talent, and approach every day like it's Day 1. We're a globally diverse team with hubs in New York City, Mountain View, Latin America, and India-embracing both hybrid and remote work to bring the best minds together, wherever they are. If you're driven by continuous learning, rapid pivots, and the challenges of building in a high-growth startup, we'd love to talk. This is more than a job-it's a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns the entire infrastructure stacks. SREs design and implement the tools that automate building reliable and performant systems. We emphasize building tools over manual processes. We implement, not administer. We're obsessed with automation, not repetition. Our job is to focus on building reliable infrastructure and tools for our product teams so that they can solve customer problems and deliver new features, not reinvent platforms. What you'll do Work with product engineering teams on service architecture and implementation Deliver Infrastructure configuration as code and automate everything Direct and implement monitoring and alerting systems to support rapid problem diagnosis Perform Root Cause Analysis and design and deliver resolutions Work on our Kubernetes / AWS infrastructure to support our product engineers Design secure and performant networking solutions in our production systems What you'll need +4 years of relevant experience bringing software to production at high scale Participation in on-call rotation, triaging and addressing production issues Obsession with automation and instrumentation Understanding of complex systems and failure scenarios Excellent communication skills Knowledge of AWS services, containers and container management frameworks Familiarity with Message Bus based systems and distributed architectures Proficiency in Terraform , Python and/or Go What we'd like to see BS or MS degree in the Computer Science field, or equivalent hands-on experience. Experience in product oriented environments Scalable distributed applications experience Benefits Competitive compensation with stock options Comprehensive medical, vision, and dental insurance 401k matching Fitness and wellness stipend Mobile phone reimbursement Mental well-being benefits Professional learning and development stipend Parental leave, including adoptive and foster parents 3 weeks paid time off (increases with tenure) and unlimited sick leave Compensation is a combination of salary and performance bonus Separately, you will also receive an equity grant ASAPP is committed to creating a diverse environment and is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, disability, age, or veteran status. If you have a disability and need assistance with our employment application process, please email us at to obtain assistance. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
08/05/2026
Full time
Job Description Job Description At ASAPP, our mission is simple: deliver the best AI-powered customer experience-faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed, ownership, and a relentless focus on outcomes. We work in tight, skilled teams, prioritize clarity over complexity, and continuously evolve through curiosity, data, and craftsmanship. We're seeking technologists and problem solvers who thrive in fast-paced environments, love collaborating with great talent, and approach every day like it's Day 1. We're a globally diverse team with hubs in New York City, Mountain View, Latin America, and India-embracing both hybrid and remote work to bring the best minds together, wherever they are. If you're driven by continuous learning, rapid pivots, and the challenges of building in a high-growth startup, we'd love to talk. This is more than a job-it's a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns the entire infrastructure stacks. SREs design and implement the tools that automate building reliable and performant systems. We emphasize building tools over manual processes. We implement, not administer. We're obsessed with automation, not repetition. Our job is to focus on building reliable infrastructure and tools for our product teams so that they can solve customer problems and deliver new features, not reinvent platforms. What you'll do Work with product engineering teams on service architecture and implementation Deliver Infrastructure configuration as code and automate everything Direct and implement monitoring and alerting systems to support rapid problem diagnosis Perform Root Cause Analysis and design and deliver resolutions Work on our Kubernetes / AWS infrastructure to support our product engineers Design secure and performant networking solutions in our production systems What you'll need +4 years of relevant experience bringing software to production at high scale Participation in on-call rotation, triaging and addressing production issues Obsession with automation and instrumentation Understanding of complex systems and failure scenarios Excellent communication skills Knowledge of AWS services, containers and container management frameworks Familiarity with Message Bus based systems and distributed architectures Proficiency in Terraform , Python and/or Go What we'd like to see BS or MS degree in the Computer Science field, or equivalent hands-on experience. Experience in product oriented environments Scalable distributed applications experience Benefits Competitive compensation with stock options Comprehensive medical, vision, and dental insurance 401k matching Fitness and wellness stipend Mobile phone reimbursement Mental well-being benefits Professional learning and development stipend Parental leave, including adoptive and foster parents 3 weeks paid time off (increases with tenure) and unlimited sick leave Compensation is a combination of salary and performance bonus Separately, you will also receive an equity grant ASAPP is committed to creating a diverse environment and is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, disability, age, or veteran status. If you have a disability and need assistance with our employment application process, please email us at to obtain assistance. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Senior Engineer, Performance Architecture
Samsung Semiconductor San Jose, California
Job Description Job Description Please Note: To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period. Advancing the World's Technology Together Our technology solutions power the tools you use every day including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you'll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what's possible and powering the future. We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We're dedicated to empowering people to be their true selves. Together, we're building a better tomorrow for our employees, customers, partners, and communities. The AGI (Artificial General Intelligence) Computing Lab is dedicated to solving the complex system-level challenges posed by the growing demands of future AI/ML workloads. Our team is committed to designing and developing scalable platforms that can effectively handle the computational and memory requirements of these workloads while minimizing energy consumption and maximizing performance. To achieve this goal, we collaborate closely with both hardware and software engineers to identify and address the unique challenges posed by AI/ML workloads and to explore new computing abstractions that can provide a better balance between the hardware and software components of our systems. Additionally, we continuously conduct research and development in emerging technologies and trends across memory, computing, interconnect, and AI/ML, ensuring that our platforms are always equipped to handle the most demanding workloads of the future. By working together as a dedicated and passionate team, we aim to revolutionize the way AI/ML applications are deployed and executed, ultimately contributing to the advancement of AGI in an affordable and sustainable manner. Join us in our passion to shape the future of computing! We are looking for a Senior Engineer, Performance Architecture. This role is being offered under the AGICL lab as a part of DSRA. We are a research-driven systems lab working at the intersection of large language models, accelerator hardware, and high-performance software stacks. Our mission is to design, prototype, and optimize next-generation AI systems through tight hardware-software co-design. Location: Daily onsite presence at our San Jose, CA office / U.S. headquarters in alignment with our Flexible Work policy. What You'll Do Model the architecture, performance, and power characteristics of Memory Centric Computing platforms Develop and optimize models to explore a large design space. Analyze various trade-offs within a design space, considering different architectural choices and workloads. Collaborate with architecture, design and software engineers to ensure that the PPA of modeled systems meet the requirements of our users Conduct research and development in emerging technologies and trends in AI/ML workloads and Memory Centric Computing architectures Communicate effectively with stakeholders, including users, partners, and management, to ensure that the systems are delivered on time and within budget What You Bring BS in Computer/Electrical Engineering or Computer Science with 5+ years of working experiences in silicon development or MS in Computer/Electrical Engineering or Computer Science with 3+ years of relevant working experience or PhD and 0+ years of relevant working experience preferred. Strong background in computer architecture Experiences in developing and optimizing models for high-performance computing systems Strong analytical and problem-solving skills Excellent communication and interpersonal skills Ability to work independently and as part of a team You're inclusive, adapting your style to the situation and diverse global norms of our people. An avid learner, you approach challenges with curiosity and resilience, seeking data to help build understanding. You're collaborative, building relationships, humbly offering support and openly welcoming approaches. Innovative and creative, you proactively explore new ideas and adapt quickly to change. What We Offer The pay range below is for all roles at this level across all US locations and functions. Pay within this range varies by work location and may also depend on job-related knowledge, skills, and experience. We also offer incentive opportunities that reward employees based on individual and company performance. This is in addition to our diverse package of benefits centered around the wellbeing of our employees and their loved ones. In addition to the usual Medical/Dental/Vision/401k, our inclusive rewards plan empowers our people to care for their whole selves. An investment in your future is an investment in ours. Give Back With a charitable giving match and frequent opportunities to get involved, we take an active role in supporting the community. Enjoy Time Away You'll start with 4+ weeks of paid time off a year, plus holidays and sick leave, to rest and recharge. Care for Family Whatever family means to you, we want to support you along the way-including a stipend for fertility care or adoption, medical travel support, and virtual vet care for your fur babies. Prioritize Emotional Wellness With on-demand apps and free confidential therapy sessions, you'll have support no matter where you are. Stay Fit Eating well and being active are important parts of a healthy life. Our onsite Café and gym, plus virtual classes, make it easier. Embrace Flexibility Benefits are best when you have the space to use them. That's why we facilitate a flexible environment so you can find the right balance for you. Base Pay Range $138,000-$206,000 USD Equal Opportunity Employment Policy Samsung Semiconductor takes pride in being an equal opportunity workplace dedicated to fostering an environment where all individuals feel valued and empowered to excel, regardless of race, religion, color, age, disability, sex, gender identity, sexual orientation, ancestry, genetic information, marital status, national origin, political affiliation, or veteran status. When selecting team members, we prioritize talent and qualities such as humility, kindness, and dedication. We extend comprehensive accommodations throughout our recruiting processes for candidates with disabilities, long-term conditions, neurodivergent individuals, or those requiring pregnancy-related support. All candidates scheduled for an interview will receive guidance on requesting accommodations. Our Commitment to Innovation and Fairness At Samsung Semiconductor, we use Artificial Intelligence (AI) tools in the recruitment process to enhance efficiency. However, AI is used as a support tool, not a final decision-maker. All hiring decisions are made by our human recruiting team and hiring managers to ensure every candidate is evaluated fairly and holistically. Recruiting Agency Policy We do not accept unsolicited resumes. Only authorized recruitment agencies that have a current and valid agreement with Samsung Semiconductor, Inc. are permitted to submit resumes for any job openings. Applicant AI Use Policy At Samsung Semiconductor, we support innovation and technology. However, to ensure a fair and authentic assessment, we ask that candidates rely on their own knowledge and skills throughout the process. AI tools may be used for basic preparation, grammar, and research, but should not be used to generate or assist with submitted content or live interview responses. If we determine that AI is being used outside these guidelines, we reserve the right to pause or end the interview, and your candidacy may be disqualified. Trade Secret Notice By submitting an application, you agree not to disclose to Samsung-or encourage Samsung to use-any confidential or proprietary information (including trade secrets) belonging to a current or former employer or other entity. Applicant Privacy Policy - us/careers/us/privacy/
08/05/2026
Full time
Job Description Job Description Please Note: To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period. Advancing the World's Technology Together Our technology solutions power the tools you use every day including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you'll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what's possible and powering the future. We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We're dedicated to empowering people to be their true selves. Together, we're building a better tomorrow for our employees, customers, partners, and communities. The AGI (Artificial General Intelligence) Computing Lab is dedicated to solving the complex system-level challenges posed by the growing demands of future AI/ML workloads. Our team is committed to designing and developing scalable platforms that can effectively handle the computational and memory requirements of these workloads while minimizing energy consumption and maximizing performance. To achieve this goal, we collaborate closely with both hardware and software engineers to identify and address the unique challenges posed by AI/ML workloads and to explore new computing abstractions that can provide a better balance between the hardware and software components of our systems. Additionally, we continuously conduct research and development in emerging technologies and trends across memory, computing, interconnect, and AI/ML, ensuring that our platforms are always equipped to handle the most demanding workloads of the future. By working together as a dedicated and passionate team, we aim to revolutionize the way AI/ML applications are deployed and executed, ultimately contributing to the advancement of AGI in an affordable and sustainable manner. Join us in our passion to shape the future of computing! We are looking for a Senior Engineer, Performance Architecture. This role is being offered under the AGICL lab as a part of DSRA. We are a research-driven systems lab working at the intersection of large language models, accelerator hardware, and high-performance software stacks. Our mission is to design, prototype, and optimize next-generation AI systems through tight hardware-software co-design. Location: Daily onsite presence at our San Jose, CA office / U.S. headquarters in alignment with our Flexible Work policy. What You'll Do Model the architecture, performance, and power characteristics of Memory Centric Computing platforms Develop and optimize models to explore a large design space. Analyze various trade-offs within a design space, considering different architectural choices and workloads. Collaborate with architecture, design and software engineers to ensure that the PPA of modeled systems meet the requirements of our users Conduct research and development in emerging technologies and trends in AI/ML workloads and Memory Centric Computing architectures Communicate effectively with stakeholders, including users, partners, and management, to ensure that the systems are delivered on time and within budget What You Bring BS in Computer/Electrical Engineering or Computer Science with 5+ years of working experiences in silicon development or MS in Computer/Electrical Engineering or Computer Science with 3+ years of relevant working experience or PhD and 0+ years of relevant working experience preferred. Strong background in computer architecture Experiences in developing and optimizing models for high-performance computing systems Strong analytical and problem-solving skills Excellent communication and interpersonal skills Ability to work independently and as part of a team You're inclusive, adapting your style to the situation and diverse global norms of our people. An avid learner, you approach challenges with curiosity and resilience, seeking data to help build understanding. You're collaborative, building relationships, humbly offering support and openly welcoming approaches. Innovative and creative, you proactively explore new ideas and adapt quickly to change. What We Offer The pay range below is for all roles at this level across all US locations and functions. Pay within this range varies by work location and may also depend on job-related knowledge, skills, and experience. We also offer incentive opportunities that reward employees based on individual and company performance. This is in addition to our diverse package of benefits centered around the wellbeing of our employees and their loved ones. In addition to the usual Medical/Dental/Vision/401k, our inclusive rewards plan empowers our people to care for their whole selves. An investment in your future is an investment in ours. Give Back With a charitable giving match and frequent opportunities to get involved, we take an active role in supporting the community. Enjoy Time Away You'll start with 4+ weeks of paid time off a year, plus holidays and sick leave, to rest and recharge. Care for Family Whatever family means to you, we want to support you along the way-including a stipend for fertility care or adoption, medical travel support, and virtual vet care for your fur babies. Prioritize Emotional Wellness With on-demand apps and free confidential therapy sessions, you'll have support no matter where you are. Stay Fit Eating well and being active are important parts of a healthy life. Our onsite Café and gym, plus virtual classes, make it easier. Embrace Flexibility Benefits are best when you have the space to use them. That's why we facilitate a flexible environment so you can find the right balance for you. Base Pay Range $138,000-$206,000 USD Equal Opportunity Employment Policy Samsung Semiconductor takes pride in being an equal opportunity workplace dedicated to fostering an environment where all individuals feel valued and empowered to excel, regardless of race, religion, color, age, disability, sex, gender identity, sexual orientation, ancestry, genetic information, marital status, national origin, political affiliation, or veteran status. When selecting team members, we prioritize talent and qualities such as humility, kindness, and dedication. We extend comprehensive accommodations throughout our recruiting processes for candidates with disabilities, long-term conditions, neurodivergent individuals, or those requiring pregnancy-related support. All candidates scheduled for an interview will receive guidance on requesting accommodations. Our Commitment to Innovation and Fairness At Samsung Semiconductor, we use Artificial Intelligence (AI) tools in the recruitment process to enhance efficiency. However, AI is used as a support tool, not a final decision-maker. All hiring decisions are made by our human recruiting team and hiring managers to ensure every candidate is evaluated fairly and holistically. Recruiting Agency Policy We do not accept unsolicited resumes. Only authorized recruitment agencies that have a current and valid agreement with Samsung Semiconductor, Inc. are permitted to submit resumes for any job openings. Applicant AI Use Policy At Samsung Semiconductor, we support innovation and technology. However, to ensure a fair and authentic assessment, we ask that candidates rely on their own knowledge and skills throughout the process. AI tools may be used for basic preparation, grammar, and research, but should not be used to generate or assist with submitted content or live interview responses. If we determine that AI is being used outside these guidelines, we reserve the right to pause or end the interview, and your candidacy may be disqualified. Trade Secret Notice By submitting an application, you agree not to disclose to Samsung-or encourage Samsung to use-any confidential or proprietary information (including trade secrets) belonging to a current or former employer or other entity. Applicant Privacy Policy - us/careers/us/privacy/
Site Reliability Engineer (SRE) / DevOps Engineer
E-Space Saratoga, California
Job Description Job Description Ready to make connectivity from space universally accessible, secure and actionable? Then you've come to the right place! E-Space is bridging Earth and space to enable hyper-scaled deployments of Internet of Things (IoT) solutions and services. We are building a highly-advanced low Earth orbit (LEO) space system that will fundamentally change the design, economics, manufacturing and service delivery associated with traditional satellite and terrestrial IoT systems. We're intentional, we're unapologetically curious and we're 100% committed to innovate space-based communications and deliver actionable intelligence that will expand global economies, protect space and our planet and enhance our overall quality of life. We are seeking a DevOps Engineer who is eager to have an immediate impact in establishing and building out our development, testing, and release platforms, taking us from square one to a sophisticated DevOps environment. You will be working with key engineering stakeholders across the company to ensure we are following test-and-release best practices, establishing best-in-class technology and tool stacks, and keeping our environments safe and secure from outside threats. What you will be doing: Design, deploy, and maintain highly-scalable, highly-available software systems in AWS Architect and manage containerized applications on Amazon EKS with focus on reliability and performance Build and maintain Infrastructure as Code using Terraform for AWS cloud resources Develop and optimize CI/CD pipelines for automated testing, deployment, and rollback capabilities Implement comprehensive monitoring, alerting, and observability solutions using CloudWatch, Prometheus, and Grafana Ensure system reliability through SLI/SLO definition, error budgets, and incident response procedures Collaborate directly with engineering teams to optimize application deployment and operations Manage deployments and scaling strategies to support mission-critical operations Automate and enforce cloud security, governance, and compliance controls Participate in on-call rotation and lead incident response for production level systems What you bring to this role: 5+ years of experience in SRE, DevOps, or Platform Engineering roles Proven experience designing and operating mission-critical, highly-available systems within AWS Advanced proficiency in Infrastructure as Code using Terraform (OpenTofu) Deep experience with Kubernetes, EKS, Helm, and container orchestration Strong CI/CD pipeline development and management experience (Bitbucket preferred) Proficiency in Python and Bash scripting for automation Experience with monitoring and observability tools (Prometheus, Grafana, ELK Stack) Knowledge of capacity planning and performance optimization Experience with database operations and scaling (RDS, Aurora, or similar) Extra bonus points for the following: AWS Solutions Architect Professional, Certified Kubernetes Administrator (CKA), or equivalent expertise Experience with incident management and post-mortem processes Experience with GitOps workflows and tools (ArgoCD, Flux) Knowledge of service mesh technologies (Istio, Linkerd) Experience with chaos engineering and disaster recovery planning Experience with Zero Trust Networking (ZTNA) or VPN solutions Background in aerospace, defense, or other mission-critical industries Strong intellectual curiosity and commitment to continuous learning Exceptional attention to detail and an ownership mentality The estimated range is meant to reflect an anticipated salary range for the position in question, which is based on market data and other factors, all of which are subject to change. Individual pay is based on location, skills and expertise, depth of relevant experience, and other relevant factors. For questions about this, please speak to the recruiter if you decide to apply for the role and are selected for an interview . This is a full time, exempt position, based out of our Saratoga office. The target base pay for this position is $100,000 - $170,000 annually. The total compensation packaged will be determined by various factors such as your relevant job-related knowledge, skills, and experience. We are redefining how satellites are designed, manufactured and used-so we're looking for candidates with passion, deep knowledge and direct experience on LEO satellite component development, design and in-orbit activities. If that's your experience - then we'll be immediately wow-ed. E-Space is not currently able to provide employment sponsorship for candidates who do not hold work authorization for the location of this role. Why E-Space is right for you: As a member of our team, you will play a crucial role in driving our success. Our team members have a strong sense of dedication and responsibility; this includes a strong commitment to our mission to create an entirely new suite of global capabilities to improve lives, business efficiencies and build a smarter planet. This means that there will be times when extra hours, including nights and weekends, may be needed to meet critical deadlines and mission goals. In return, we offer a dynamic work environment with opportunities for professional growth and development and the chance to make a meaningful impact in a high-growth industry. We want you to make the most of your journey at E-Space. That's why we support and invest in the physical, emotional and financial well-being of our team members and their families. Some of what you can expect when working at E-Space: • An opportunity to really make a difference • Sustainability at our core • Fair and honest workplace • Innovative thinking is encouraged • Competitive salaries • Continuous learning and development • Health and wellness care options • Financial solutions for the future • Optional legal services (US only) • Paid holidays • Paid time off We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
08/05/2026
Full time
Job Description Job Description Ready to make connectivity from space universally accessible, secure and actionable? Then you've come to the right place! E-Space is bridging Earth and space to enable hyper-scaled deployments of Internet of Things (IoT) solutions and services. We are building a highly-advanced low Earth orbit (LEO) space system that will fundamentally change the design, economics, manufacturing and service delivery associated with traditional satellite and terrestrial IoT systems. We're intentional, we're unapologetically curious and we're 100% committed to innovate space-based communications and deliver actionable intelligence that will expand global economies, protect space and our planet and enhance our overall quality of life. We are seeking a DevOps Engineer who is eager to have an immediate impact in establishing and building out our development, testing, and release platforms, taking us from square one to a sophisticated DevOps environment. You will be working with key engineering stakeholders across the company to ensure we are following test-and-release best practices, establishing best-in-class technology and tool stacks, and keeping our environments safe and secure from outside threats. What you will be doing: Design, deploy, and maintain highly-scalable, highly-available software systems in AWS Architect and manage containerized applications on Amazon EKS with focus on reliability and performance Build and maintain Infrastructure as Code using Terraform for AWS cloud resources Develop and optimize CI/CD pipelines for automated testing, deployment, and rollback capabilities Implement comprehensive monitoring, alerting, and observability solutions using CloudWatch, Prometheus, and Grafana Ensure system reliability through SLI/SLO definition, error budgets, and incident response procedures Collaborate directly with engineering teams to optimize application deployment and operations Manage deployments and scaling strategies to support mission-critical operations Automate and enforce cloud security, governance, and compliance controls Participate in on-call rotation and lead incident response for production level systems What you bring to this role: 5+ years of experience in SRE, DevOps, or Platform Engineering roles Proven experience designing and operating mission-critical, highly-available systems within AWS Advanced proficiency in Infrastructure as Code using Terraform (OpenTofu) Deep experience with Kubernetes, EKS, Helm, and container orchestration Strong CI/CD pipeline development and management experience (Bitbucket preferred) Proficiency in Python and Bash scripting for automation Experience with monitoring and observability tools (Prometheus, Grafana, ELK Stack) Knowledge of capacity planning and performance optimization Experience with database operations and scaling (RDS, Aurora, or similar) Extra bonus points for the following: AWS Solutions Architect Professional, Certified Kubernetes Administrator (CKA), or equivalent expertise Experience with incident management and post-mortem processes Experience with GitOps workflows and tools (ArgoCD, Flux) Knowledge of service mesh technologies (Istio, Linkerd) Experience with chaos engineering and disaster recovery planning Experience with Zero Trust Networking (ZTNA) or VPN solutions Background in aerospace, defense, or other mission-critical industries Strong intellectual curiosity and commitment to continuous learning Exceptional attention to detail and an ownership mentality The estimated range is meant to reflect an anticipated salary range for the position in question, which is based on market data and other factors, all of which are subject to change. Individual pay is based on location, skills and expertise, depth of relevant experience, and other relevant factors. For questions about this, please speak to the recruiter if you decide to apply for the role and are selected for an interview . This is a full time, exempt position, based out of our Saratoga office. The target base pay for this position is $100,000 - $170,000 annually. The total compensation packaged will be determined by various factors such as your relevant job-related knowledge, skills, and experience. We are redefining how satellites are designed, manufactured and used-so we're looking for candidates with passion, deep knowledge and direct experience on LEO satellite component development, design and in-orbit activities. If that's your experience - then we'll be immediately wow-ed. E-Space is not currently able to provide employment sponsorship for candidates who do not hold work authorization for the location of this role. Why E-Space is right for you: As a member of our team, you will play a crucial role in driving our success. Our team members have a strong sense of dedication and responsibility; this includes a strong commitment to our mission to create an entirely new suite of global capabilities to improve lives, business efficiencies and build a smarter planet. This means that there will be times when extra hours, including nights and weekends, may be needed to meet critical deadlines and mission goals. In return, we offer a dynamic work environment with opportunities for professional growth and development and the chance to make a meaningful impact in a high-growth industry. We want you to make the most of your journey at E-Space. That's why we support and invest in the physical, emotional and financial well-being of our team members and their families. Some of what you can expect when working at E-Space: • An opportunity to really make a difference • Sustainability at our core • Fair and honest workplace • Innovative thinking is encouraged • Competitive salaries • Continuous learning and development • Health and wellness care options • Financial solutions for the future • Optional legal services (US only) • Paid holidays • Paid time off We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Senior Device Engineer - Lifecycle Management
Allergan Aesthetics Pleasanton, California
Job Description Job Description Company Description About AbbVie AbbVie's mission is to discover and deliver innovative medicines and solutions that solve serious health issues today and address the medical challenges of tomorrow. We strive to have a remarkable impact on people's lives across several key therapeutic areas including immunology, oncology and neuroscience - and products and services in our Allergan Aesthetics portfolio. For more information about AbbVie, please visit us at . Follow us on LinkedIn, Facebook, Instagram, X and YouTube. Job Description The Senior Software Engineer - Lifecycle Management will work collaboratively with a team to support medical device products from the transfer to production through product end of life. The Engineer will have a technical leadership role for supporting components and subassemblies of the body contouring products and will closely interact with a multi-disciplined Engineering team consisting of electrical, software and mechanical groups. The individual works within cross-functional teams and provides software requirements, design and implementation for current generation software and systems projects. He or she develops a thorough understanding of design requirements to ensure that the system's objectives are properly defined and ultimately achieved. This role is focused on continuous improvement of existing products. This individual must have strong technical skills complemented by great communications and teamwork qualities. Experience in a software development background in a structured/regulated environment such as medical device development is required. Responsibilities Lead and manage small scale projects for on time deliverable. Contribute to requirements definition at the functional level and work with cross functional groups. Perform in-depth data analysis and drive improvements to software or product quality. Design, develop, and support embedded, Windows embedded and desktop applications. Participate in software work product reviews/inspections. Interface, integrate, troubleshoot and debug software and hardware components. Generate required product development documentation including functional specifications and design documents. Execute manual or automated tests for verification and validation of software applications. Design, code and validate software tools for use in the verification and manufacturing of the product. Work with Software Verification, Product Support and Manufacturing to resolve software issues. Drive improvements to process quality. Responsible for performing all duties in compliance with FDA's Quality System Regulation (QSR), ISO13485, the Canadian Medical Device Regulations, and all other international regulatory requirements with which AbbVie complies. Qualifications BS in Software Engineering or equivalent degree and/or experience. Advanced degree desirable. Minimum of 8+ years experience in engineering design and at least 5 years of experience with embedded Windows programming with C# and . NET. At least 3 years of experience in medical devices or similarly controlled software environment preferred. Experience in developing event driven, multi-threaded Windows-based applications using .NET Framework and C# preferred. Required experience in structured software and systems development and integration, including experience in software design methodologies, design patterns, component-oriented software architecture to produce high-quality software applications. Experience with common protocols: RS232, SPI, USB a plus. Knowledge of PID control algorithm. Knowledge of software life cycle processes used in regulated development environments such as IEC 62304. Result-oriented, self-motivated and able to participate as both a team member and an individual contributor. Proficiency in MS Office, including Word and Excel. Additional Information Applicable only to applicants applying to a position in any location with pay disclosure requirements under state or local law: The compensation range described below is the range of possible base pay compensation that the Company believes in good faith it will pay for this role at the time of this posting based on the job grade for this position. Individual compensation paid within this range will depend on many factors including geographic location, and we may ultimately pay more or less than the posted range. This range may be modified in the future. We offer a comprehensive package of benefits including paid time off (vacation, holidays, sick), medical/dental/vision insurance and 401(k) to eligible employees. This job is eligible to participate in our long-term incentive programs. Note: No amount of pay is considered to be wages or compensation until such amount is earned, vested, and determinable. The amount and availability of any bonus, commission, incentive, benefits, or any other form of compensation and benefits that are allocable to a particular employee remains in the Company's sole and absolute discretion unless and until paid and may be modified at the Company's sole and absolute discretion, consistent with applicable law. AbbVie is an equal opportunity employer and is committed to operating with integrity, driving innovation, transforming lives and serving our community. Equal Opportunity Employer/Veterans/Disabled. US & Puerto Rico only - to learn more, visit -us/equal-employment-opportunity-employer.html US & Puerto Rico applicants seeking a reasonable accommodation, click here to learn more: -us/reasonable- accommodations.html
08/05/2026
Full time
Job Description Job Description Company Description About AbbVie AbbVie's mission is to discover and deliver innovative medicines and solutions that solve serious health issues today and address the medical challenges of tomorrow. We strive to have a remarkable impact on people's lives across several key therapeutic areas including immunology, oncology and neuroscience - and products and services in our Allergan Aesthetics portfolio. For more information about AbbVie, please visit us at . Follow us on LinkedIn, Facebook, Instagram, X and YouTube. Job Description The Senior Software Engineer - Lifecycle Management will work collaboratively with a team to support medical device products from the transfer to production through product end of life. The Engineer will have a technical leadership role for supporting components and subassemblies of the body contouring products and will closely interact with a multi-disciplined Engineering team consisting of electrical, software and mechanical groups. The individual works within cross-functional teams and provides software requirements, design and implementation for current generation software and systems projects. He or she develops a thorough understanding of design requirements to ensure that the system's objectives are properly defined and ultimately achieved. This role is focused on continuous improvement of existing products. This individual must have strong technical skills complemented by great communications and teamwork qualities. Experience in a software development background in a structured/regulated environment such as medical device development is required. Responsibilities Lead and manage small scale projects for on time deliverable. Contribute to requirements definition at the functional level and work with cross functional groups. Perform in-depth data analysis and drive improvements to software or product quality. Design, develop, and support embedded, Windows embedded and desktop applications. Participate in software work product reviews/inspections. Interface, integrate, troubleshoot and debug software and hardware components. Generate required product development documentation including functional specifications and design documents. Execute manual or automated tests for verification and validation of software applications. Design, code and validate software tools for use in the verification and manufacturing of the product. Work with Software Verification, Product Support and Manufacturing to resolve software issues. Drive improvements to process quality. Responsible for performing all duties in compliance with FDA's Quality System Regulation (QSR), ISO13485, the Canadian Medical Device Regulations, and all other international regulatory requirements with which AbbVie complies. Qualifications BS in Software Engineering or equivalent degree and/or experience. Advanced degree desirable. Minimum of 8+ years experience in engineering design and at least 5 years of experience with embedded Windows programming with C# and . NET. At least 3 years of experience in medical devices or similarly controlled software environment preferred. Experience in developing event driven, multi-threaded Windows-based applications using .NET Framework and C# preferred. Required experience in structured software and systems development and integration, including experience in software design methodologies, design patterns, component-oriented software architecture to produce high-quality software applications. Experience with common protocols: RS232, SPI, USB a plus. Knowledge of PID control algorithm. Knowledge of software life cycle processes used in regulated development environments such as IEC 62304. Result-oriented, self-motivated and able to participate as both a team member and an individual contributor. Proficiency in MS Office, including Word and Excel. Additional Information Applicable only to applicants applying to a position in any location with pay disclosure requirements under state or local law: The compensation range described below is the range of possible base pay compensation that the Company believes in good faith it will pay for this role at the time of this posting based on the job grade for this position. Individual compensation paid within this range will depend on many factors including geographic location, and we may ultimately pay more or less than the posted range. This range may be modified in the future. We offer a comprehensive package of benefits including paid time off (vacation, holidays, sick), medical/dental/vision insurance and 401(k) to eligible employees. This job is eligible to participate in our long-term incentive programs. Note: No amount of pay is considered to be wages or compensation until such amount is earned, vested, and determinable. The amount and availability of any bonus, commission, incentive, benefits, or any other form of compensation and benefits that are allocable to a particular employee remains in the Company's sole and absolute discretion unless and until paid and may be modified at the Company's sole and absolute discretion, consistent with applicable law. AbbVie is an equal opportunity employer and is committed to operating with integrity, driving innovation, transforming lives and serving our community. Equal Opportunity Employer/Veterans/Disabled. US & Puerto Rico only - to learn more, visit -us/equal-employment-opportunity-employer.html US & Puerto Rico applicants seeking a reasonable accommodation, click here to learn more: -us/reasonable- accommodations.html
Senior Site Reliability Engineer- Palo Alto, the US
Kody Palo Alto, California
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Cross-Border Collaboration: Act as a key technical bridge between our US operations and international engineering hubs, leveraging bilingual communication to streamline complex technical alignment. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
08/05/2026
Full time
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Cross-Border Collaboration: Act as a key technical bridge between our US operations and international engineering hubs, leveraging bilingual communication to streamline complex technical alignment. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
Senior Site Reliability Engineer- Sunnyvale, CA, the US
Kody Sunnyvale, California
Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Lead a dynamic and innovative team in a very rapidly growing company. Competitive package. Collaborative, inclusive environment where your contributions are recognized and valued.
08/05/2026
Full time
Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Lead a dynamic and innovative team in a very rapidly growing company. Competitive package. Collaborative, inclusive environment where your contributions are recognized and valued.
Site Reliability Engineer
VantageScore San Francisco, California
Job Description Job Description About The Role We are seeking an experienced Site Reliability Engineer (SRE) with a strong focus on DevSecOps to join our growing engineering team. In this role, you will oversee and maintain the reliability, security posture, and operational hygiene of our cloud infrastructure, APIs, and software supply chain. You will drive patch management programs, harden our Cloud infrastructure, and maintain our code repositories to ensure all systems remain compliant, secure, and scalable. This role is ideal for an engineer who thrives at the intersection of operations and security, is passionate about automation, and takes pride in keeping complex environments clean, auditable, and resilient. Key Responsibilities Own and execute end-to-end patch management across AWS compute resources (EC2, ECS, Lambda runtimes, EKS nodes), third-party dependencies, and OS-level packages. Monitor, triage, and remediate vulnerabilities identified by security scanning tools (e.g., AWS Inspector, Dependabot, Security Hub, or equivalent), prioritizing by CVSS severity and business impact. Maintain and enforce branch protection rules, secret scanning policies, and dependency update workflows across all code repositories. Design and implement automated pipelines for continuous compliance checking, security testing (SAST/DAST/SCA), and infrastructure drift detection. Collaborate with IT & Info-Sec SMEs on AWS IAM roles and policies, VPC configurations, Security Groups, CloudTrail, Config, and GuardDuty to ensure least-privilege access and auditability. Collaborate with development teams to embed security controls into CI/CD pipelines (GitHub Actions, CodePipeline, or equivalent) without impeding developer velocity. Support the reliability and availability of production APIs - including uptime monitoring, incident response, runbook creation, and post-incident reviews. Partner with Legal and Data Governance SMEs on API access procedures and monitoring. Define and track SLOs/SLAs for internal and external APIs; implement alerting and dashboards using observability tooling (e.g., CloudWatch, Datadog, Grafana). Lead periodic infrastructure and dependency audits; produce clear reports on patch compliance status and open risk items for engineering and security leadership. Maintain thorough documentation of patching schedules, runbooks, access policies, and environment configurations. Participate in on-call rotation and contribute to a culture of continuous improvement. Required Qualifications Bachelor's Degree in Computer Science, Information Systems, or a related field (or equivalent practical experience). 5+ years of professional experience in a Site Reliability Engineering, Software Engineering, DevOps, or DevSecOps role. Demonstrated expertise managing AWS environments - including EC2, Lambda, ECS/EKS, S3, RDS, IAM, VPC, CloudTrail, Config, and GuardDuty. Experience with various cloud environments: AWS, Azure, GPC Strong experience with GitHub administration: branch protection, Actions workflows, secret scanning, Dependabot, and code owners. Hands-on experience with patch management and vulnerability remediation at scale, including OS-level patching (Amazon Linux, Ubuntu) and dependency lifecycle management. Proficiency with infrastructure-as-code tools (Terraform, CloudFormation, or AWS CDK). Experience integrating security tooling (SAST, DAST, SCA, container scanning) into CI/CD pipelines. Solid understanding of API reliability patterns: health checks, rate limiting, circuit breakers, and observability. Familiarity with compliance frameworks relevant to cloud environments (SOC 2, CIS Benchmarks, NIST CSF). Strong scripting skills in Python, Bash, or similar for automation and tooling. Excellent communication skills and ability to translate technical risk for non-technical stakeholders. Build observation (logging, metrics, alerting) systems to make sure system works well, and develop response plans. Preferred Qualifications AWS certifications (e.g., AWS Certified Security - Specialty, AWS Certified DevOps Engineer - Professional). Experience with container security and Kubernetes (EKS) hardening. Familiarity with CSPM tools (e.g., Wiz, Prisma Cloud, AWS Security Hub) for continuous cloud posture management. Experience managing API gateways (AWS API Gateway, Kong, or similar) including security policy enforcement. Exposure to secrets management solutions (AWS Secrets Manager, HashiCorp Vault). Knowledge of SBOM (Software Bill of Materials) generation and management. Experience with incident response playbooks and tabletop exercises. Familiarity with Agile/Scrum methodologies and cross-functional engineering teams. Compensation The anticipated base salary range for this position is $150,000 annually, plus eligibility for a 15% annual performance bonus. Actual compensation will be determined based on several factors, including skills, experience, education, certifications, and geographic location. In addition to base salary and bonus eligibility, we offer a competitive benefits package, including medical, dental, vision, 401(k), paid time off, and other employee benefits.
08/05/2026
Full time
Job Description Job Description About The Role We are seeking an experienced Site Reliability Engineer (SRE) with a strong focus on DevSecOps to join our growing engineering team. In this role, you will oversee and maintain the reliability, security posture, and operational hygiene of our cloud infrastructure, APIs, and software supply chain. You will drive patch management programs, harden our Cloud infrastructure, and maintain our code repositories to ensure all systems remain compliant, secure, and scalable. This role is ideal for an engineer who thrives at the intersection of operations and security, is passionate about automation, and takes pride in keeping complex environments clean, auditable, and resilient. Key Responsibilities Own and execute end-to-end patch management across AWS compute resources (EC2, ECS, Lambda runtimes, EKS nodes), third-party dependencies, and OS-level packages. Monitor, triage, and remediate vulnerabilities identified by security scanning tools (e.g., AWS Inspector, Dependabot, Security Hub, or equivalent), prioritizing by CVSS severity and business impact. Maintain and enforce branch protection rules, secret scanning policies, and dependency update workflows across all code repositories. Design and implement automated pipelines for continuous compliance checking, security testing (SAST/DAST/SCA), and infrastructure drift detection. Collaborate with IT & Info-Sec SMEs on AWS IAM roles and policies, VPC configurations, Security Groups, CloudTrail, Config, and GuardDuty to ensure least-privilege access and auditability. Collaborate with development teams to embed security controls into CI/CD pipelines (GitHub Actions, CodePipeline, or equivalent) without impeding developer velocity. Support the reliability and availability of production APIs - including uptime monitoring, incident response, runbook creation, and post-incident reviews. Partner with Legal and Data Governance SMEs on API access procedures and monitoring. Define and track SLOs/SLAs for internal and external APIs; implement alerting and dashboards using observability tooling (e.g., CloudWatch, Datadog, Grafana). Lead periodic infrastructure and dependency audits; produce clear reports on patch compliance status and open risk items for engineering and security leadership. Maintain thorough documentation of patching schedules, runbooks, access policies, and environment configurations. Participate in on-call rotation and contribute to a culture of continuous improvement. Required Qualifications Bachelor's Degree in Computer Science, Information Systems, or a related field (or equivalent practical experience). 5+ years of professional experience in a Site Reliability Engineering, Software Engineering, DevOps, or DevSecOps role. Demonstrated expertise managing AWS environments - including EC2, Lambda, ECS/EKS, S3, RDS, IAM, VPC, CloudTrail, Config, and GuardDuty. Experience with various cloud environments: AWS, Azure, GPC Strong experience with GitHub administration: branch protection, Actions workflows, secret scanning, Dependabot, and code owners. Hands-on experience with patch management and vulnerability remediation at scale, including OS-level patching (Amazon Linux, Ubuntu) and dependency lifecycle management. Proficiency with infrastructure-as-code tools (Terraform, CloudFormation, or AWS CDK). Experience integrating security tooling (SAST, DAST, SCA, container scanning) into CI/CD pipelines. Solid understanding of API reliability patterns: health checks, rate limiting, circuit breakers, and observability. Familiarity with compliance frameworks relevant to cloud environments (SOC 2, CIS Benchmarks, NIST CSF). Strong scripting skills in Python, Bash, or similar for automation and tooling. Excellent communication skills and ability to translate technical risk for non-technical stakeholders. Build observation (logging, metrics, alerting) systems to make sure system works well, and develop response plans. Preferred Qualifications AWS certifications (e.g., AWS Certified Security - Specialty, AWS Certified DevOps Engineer - Professional). Experience with container security and Kubernetes (EKS) hardening. Familiarity with CSPM tools (e.g., Wiz, Prisma Cloud, AWS Security Hub) for continuous cloud posture management. Experience managing API gateways (AWS API Gateway, Kong, or similar) including security policy enforcement. Exposure to secrets management solutions (AWS Secrets Manager, HashiCorp Vault). Knowledge of SBOM (Software Bill of Materials) generation and management. Experience with incident response playbooks and tabletop exercises. Familiarity with Agile/Scrum methodologies and cross-functional engineering teams. Compensation The anticipated base salary range for this position is $150,000 annually, plus eligibility for a 15% annual performance bonus. Actual compensation will be determined based on several factors, including skills, experience, education, certifications, and geographic location. In addition to base salary and bonus eligibility, we offer a competitive benefits package, including medical, dental, vision, 401(k), paid time off, and other employee benefits.
AI Engineer
twenty80.io San Francisco, California
Job Description Job Description About This Role Twenty80 is partnering with a fast-moving technology company to find an AI Engineer to build and ship LLM-powered applications. Our client is looking for someone to own new AI workflows from prototype to production. Expect to ship fast, talk directly to customers, and build systems that solve real problems for users. This is a full-time onsite position. You will work five days a week from our client's office in San Francisco, collaborating in person with PMs, designers, and the GTM team to ship fast and stay close to the work. This is a deeply technical role that blends ML, product thinking, and engineering. You will build and own production services that rely on LLMs and speech models. What you'll do Develop and own end-to-end AI applications across a range of workflows, from copilots to automation tooling and beyond. Improve our client's LLM-powered platform, used daily by thousands of users. Build and optimize inference pipelines and real-time speech recognition systems. Collaborate closely with PMs, designers, and GTM teams to iterate rapidly and launch high-impact features. Suggested Experience Our client is looking for exceptional, top-tier talent. 3+ years building production AI systems, ideally in fast-moving environments or as a founder. Extensive experience with LLM inference (OpenAI, Anthropic, Llama, etc.). Strong product sense. Comfortable working across the stack (frontend, backend, etc.) to get features shipped. Obsessed with speed, ownership, and getting real user feedback. This is a remote position, although Bay Area residents are encouraged to work from the client's SF office. Nice to Haves Experience with ASR systems (e.g. Whisper). Seniority 2 to 6 years of experience as a software engineer, with AI and LLM-focused product work. Work experience Must have built complex AI applications end-to-end (LLM inference work, agent products, etc.). Experience in a fast-moving startup or as a founder. Education CS undergraduate degree from a four-year university. Hard skills Able to work across the stack (frontend, backend, infra) to problem solve and ship features. Applied-AI literacy. Can reason about evals, statistics, and the non-determinism of LLM systems. Familiarity with SOC-2 or sensitive data pipelines. Soft skills Obsessed with speed, ownership, and getting real user feedback. Traits to avoid Lack of progression in role after 2 to 3 years. About Twenty80 Twenty80 is an executive and specialty search firm representing this client on the search. All candidate conversations run through us.
08/05/2026
Full time
Job Description Job Description About This Role Twenty80 is partnering with a fast-moving technology company to find an AI Engineer to build and ship LLM-powered applications. Our client is looking for someone to own new AI workflows from prototype to production. Expect to ship fast, talk directly to customers, and build systems that solve real problems for users. This is a full-time onsite position. You will work five days a week from our client's office in San Francisco, collaborating in person with PMs, designers, and the GTM team to ship fast and stay close to the work. This is a deeply technical role that blends ML, product thinking, and engineering. You will build and own production services that rely on LLMs and speech models. What you'll do Develop and own end-to-end AI applications across a range of workflows, from copilots to automation tooling and beyond. Improve our client's LLM-powered platform, used daily by thousands of users. Build and optimize inference pipelines and real-time speech recognition systems. Collaborate closely with PMs, designers, and GTM teams to iterate rapidly and launch high-impact features. Suggested Experience Our client is looking for exceptional, top-tier talent. 3+ years building production AI systems, ideally in fast-moving environments or as a founder. Extensive experience with LLM inference (OpenAI, Anthropic, Llama, etc.). Strong product sense. Comfortable working across the stack (frontend, backend, etc.) to get features shipped. Obsessed with speed, ownership, and getting real user feedback. This is a remote position, although Bay Area residents are encouraged to work from the client's SF office. Nice to Haves Experience with ASR systems (e.g. Whisper). Seniority 2 to 6 years of experience as a software engineer, with AI and LLM-focused product work. Work experience Must have built complex AI applications end-to-end (LLM inference work, agent products, etc.). Experience in a fast-moving startup or as a founder. Education CS undergraduate degree from a four-year university. Hard skills Able to work across the stack (frontend, backend, infra) to problem solve and ship features. Applied-AI literacy. Can reason about evals, statistics, and the non-determinism of LLM systems. Familiarity with SOC-2 or sensitive data pipelines. Soft skills Obsessed with speed, ownership, and getting real user feedback. Traits to avoid Lack of progression in role after 2 to 3 years. About Twenty80 Twenty80 is an executive and specialty search firm representing this client on the search. All candidate conversations run through us.
Senior Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home
Ginas Tech Jobs San Francisco, California
Job Description Job Description Job Description Senior Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home As the Senior Machine Learning Engineer, you are an independent owner of critical Machine Learning (ML) subsystems in production. You take ambiguous problems, design practical solutions, and ship systems that operate reliably at scale. This is a hands-on, high-impact role focused on depth. This position is 100% Remote. Senior Machine Learning Engineer Responsibilities: - Build core Machine Learning (ML) systems that power a proactive, long-horizon Artificial Intelligence (AI) product. - Own work end-to-end: data preparation, training, evaluation, inference, and iteration. - Turn research ideas into working systems that run reliably in production. - Debug model failures and system issues using real production signals. - Iterate quickly: ship, measure outcomes, refine, and repeat. - Collaborate closely with research, product, and engineering to deliver real user impact. - Mentor and review work from other Machine Learning (ML) engineers through example and technical judgment. - Work under real production constraints: latency, cost, reliability, and safety Senior Machine Learning Engineer Outcomes: - Machine Learning (ML) models and systems in production consistently meet accuracy, latency, reliability, and efficiency targets. - Complex production issues are monitored, debugged, and resolved with minimal disruption. - Training, inference, and data pipelines are robust, scalable, and maintainable over time. - Drive measurable improvements in Machine Learning (ML) systems based on real-world signals and user feedback. - Provide mentorship and technical guidance to peers, raising the overall ML engineering standard. - Collaborate cross-functionally to ensure Machine Learning (ML) features integrate seamlessly into products and meet business goals. Qualifications Senior Machine Learning Engineer Qualifications: - Experience building and shipping Machine Learning (ML) systems used by real users. - Artificial Intelligence (AI) experience required. - Experience understanding how modern Machine Learning (ML) models behave and misbehave in production. - Experience writing strong, production-quality code and think in systems, not scripts. - Experience taking ownership, work independently, and push work across the finish line. - You learn fast, communicate clearly, and improve through iteration. - Tech Stack: GPU-based training and inference systems, JAX, Python, and PyTorch. Benefits include medical insurance, Dental, Vision, Savings Plan Options, PTO, etc. Looking to hire an Senior Machine Learning Engineer in San Francisco, CA or in other cities? Our IT recruiting agencies and staffing companies can help. We help companies that are looking to hire Senior Machine Learning Engineers for jobs in San Francisco, California and in other cities too. Please contact our IT recruiting agencies and IT staffing companies today! Additional Information Please check out all of our jobs at .
08/05/2026
Full time
Job Description Job Description Job Description Senior Machine Learning Engineer, Artificial Intelligence (AI) Required, Work From Home As the Senior Machine Learning Engineer, you are an independent owner of critical Machine Learning (ML) subsystems in production. You take ambiguous problems, design practical solutions, and ship systems that operate reliably at scale. This is a hands-on, high-impact role focused on depth. This position is 100% Remote. Senior Machine Learning Engineer Responsibilities: - Build core Machine Learning (ML) systems that power a proactive, long-horizon Artificial Intelligence (AI) product. - Own work end-to-end: data preparation, training, evaluation, inference, and iteration. - Turn research ideas into working systems that run reliably in production. - Debug model failures and system issues using real production signals. - Iterate quickly: ship, measure outcomes, refine, and repeat. - Collaborate closely with research, product, and engineering to deliver real user impact. - Mentor and review work from other Machine Learning (ML) engineers through example and technical judgment. - Work under real production constraints: latency, cost, reliability, and safety Senior Machine Learning Engineer Outcomes: - Machine Learning (ML) models and systems in production consistently meet accuracy, latency, reliability, and efficiency targets. - Complex production issues are monitored, debugged, and resolved with minimal disruption. - Training, inference, and data pipelines are robust, scalable, and maintainable over time. - Drive measurable improvements in Machine Learning (ML) systems based on real-world signals and user feedback. - Provide mentorship and technical guidance to peers, raising the overall ML engineering standard. - Collaborate cross-functionally to ensure Machine Learning (ML) features integrate seamlessly into products and meet business goals. Qualifications Senior Machine Learning Engineer Qualifications: - Experience building and shipping Machine Learning (ML) systems used by real users. - Artificial Intelligence (AI) experience required. - Experience understanding how modern Machine Learning (ML) models behave and misbehave in production. - Experience writing strong, production-quality code and think in systems, not scripts. - Experience taking ownership, work independently, and push work across the finish line. - You learn fast, communicate clearly, and improve through iteration. - Tech Stack: GPU-based training and inference systems, JAX, Python, and PyTorch. Benefits include medical insurance, Dental, Vision, Savings Plan Options, PTO, etc. Looking to hire an Senior Machine Learning Engineer in San Francisco, CA or in other cities? Our IT recruiting agencies and staffing companies can help. We help companies that are looking to hire Senior Machine Learning Engineers for jobs in San Francisco, California and in other cities too. Please contact our IT recruiting agencies and IT staffing companies today! Additional Information Please check out all of our jobs at .
Site Reliability Engineer, Client Platform
Hadrian Automation Bodega Bay, California
Job Description Job Description Hadrian - Manufacturing the Future Hadrian is building autonomous factories that help aerospace and defense companies manufacture rockets, satellites, jets, and ships up to 10x faster and up to 2x cheaper. By combining advanced software, robotics, and full-stack manufacturing, we are reinventing how America produces its most critical parts. We're accelerating our mission with the launch of Factory 3 in Mesa, Arizona, a 290,000-square-foot facility creating 350 new jobs. We are expanding rapidly to support thousands of future hires, launching Hadrian Maritime to expand into naval production, and introducing a Factory-as-a-Service model that delivers complete systems instead of individual parts. Hadrian is backed by leading investors including T. Rowe Price, Lux Capital, Founders Fund, and Andreessen Horowitz, our fast-growing team is united around reindustrializing American manufacturing for the 21st century and beyond. The Role: What You'll Do Focus on building scalable , automated solutions that ensure seamless deployments, security configurations, and efficient operational workflows for our end-users. Own , administer , and optimize MDM platforms (Fleet DM, Intune, Workspace ONE) to enforce configuration and drive self-healing by writing OS-level scripts and lightweight tools that resolve recurring user-impacting issues (disk pressure, certificate expiry, drift, broken agents) at the source instead of via tickets. Proactive Response . We want to gather telemetry data and analytics to develop an understanding of device lifecycles. We want to prevent end-user disruption by understanding when and how to act. Partner with Security, IT, and Infrastructure to translate compliance requirements (CMMC) into enforceable, code-managed baselines. Build dashboards and alerts that measure end-user experience as an SLO, not a helpdesk metric. What We're Looking For Ownership. You treat the fleet as a product, take incidents personally, and close the loop with automation rather than a runbook. Strong scripting in Python, Bash, and PowerShell. Scalability. Hands-on experience with Infrastructure as Code (IaC) and configuration management: Ansible and Terraform (or equivalents like Chef, Salt, Puppet, Pulumi). Device Management. Working knowledge of at least one major MDM (Fleet DM, Intune, Jamf, Workspace ONE) and its API surface. Security Remediation. Practical experience with patch management , vulnerability remediation , and endpoint hardening on both macOS and Windows. What Will Set You Apart Strong computer science fundamentals. You can reason about systems from the operating system up, demonstrating durable and sustainable solutions. Experience building self-healing or auto-remediation platforms (remote actions, osquery + response, custom agents). Exposure to OT (Operational Technology) systems. Comfort operating in an SRE culture: SLOs, error budgets, and blameless postmortems applied to the end-user experience. Compensation For this role, the target salary range is $164,000 - $270,000 (actual range may vary based on experience). This is the lowest to highest salary we reasonably and in good faith believe we would pay for this role at the time of this posting. We may ultimately pay more or less than the posted range, and the range may be modified in the future. An employee's pay position within the salary range will be based on several factors, including, but not limited to, relevant education, qualifications, certifications, experience, skills, geographic location, performance, and business or organizational needs. Benefits for Full-time Employees Medical, dental, vision, and life insurance plans for employees 401k Relocation support may be provided for certain situations, based on business need. Flexible vacation policy Equity ITAR Requirements To conform to U.S. Government space technology export regulations, including the International Traffic in Arms Regulations (ITAR) you must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. Hadrian Is An Equal Opportunity Employer It is the Company's policy to provide equal employment opportunity for all applicants and employees. The Company does not unlawfully discriminate on the basis of race inclusive of traits historically associated with race (including, but not limited to, hair texture and protective hairstyles, such as braids, locks and twists), color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender identity, gender expression, transgender status, national origin (including, in California, possession of a drivers license), ancestry, citizenship, age, physical or mental disability, height or weight, medical condition, family care status, military or veteran status, marital status, domestic partner status, sexual orientation, genetic information, exercise of reproductive rights, any other basis protected by local, state, or federal laws, or any combination of the above characteristics. When necessary, the Company also makes reasonable accommodations for disabled candidates and employees, including for candidates or employees who are disabled by pregnancy, childbirth, or related medical conditions.
08/05/2026
Full time
Job Description Job Description Hadrian - Manufacturing the Future Hadrian is building autonomous factories that help aerospace and defense companies manufacture rockets, satellites, jets, and ships up to 10x faster and up to 2x cheaper. By combining advanced software, robotics, and full-stack manufacturing, we are reinventing how America produces its most critical parts. We're accelerating our mission with the launch of Factory 3 in Mesa, Arizona, a 290,000-square-foot facility creating 350 new jobs. We are expanding rapidly to support thousands of future hires, launching Hadrian Maritime to expand into naval production, and introducing a Factory-as-a-Service model that delivers complete systems instead of individual parts. Hadrian is backed by leading investors including T. Rowe Price, Lux Capital, Founders Fund, and Andreessen Horowitz, our fast-growing team is united around reindustrializing American manufacturing for the 21st century and beyond. The Role: What You'll Do Focus on building scalable , automated solutions that ensure seamless deployments, security configurations, and efficient operational workflows for our end-users. Own , administer , and optimize MDM platforms (Fleet DM, Intune, Workspace ONE) to enforce configuration and drive self-healing by writing OS-level scripts and lightweight tools that resolve recurring user-impacting issues (disk pressure, certificate expiry, drift, broken agents) at the source instead of via tickets. Proactive Response . We want to gather telemetry data and analytics to develop an understanding of device lifecycles. We want to prevent end-user disruption by understanding when and how to act. Partner with Security, IT, and Infrastructure to translate compliance requirements (CMMC) into enforceable, code-managed baselines. Build dashboards and alerts that measure end-user experience as an SLO, not a helpdesk metric. What We're Looking For Ownership. You treat the fleet as a product, take incidents personally, and close the loop with automation rather than a runbook. Strong scripting in Python, Bash, and PowerShell. Scalability. Hands-on experience with Infrastructure as Code (IaC) and configuration management: Ansible and Terraform (or equivalents like Chef, Salt, Puppet, Pulumi). Device Management. Working knowledge of at least one major MDM (Fleet DM, Intune, Jamf, Workspace ONE) and its API surface. Security Remediation. Practical experience with patch management , vulnerability remediation , and endpoint hardening on both macOS and Windows. What Will Set You Apart Strong computer science fundamentals. You can reason about systems from the operating system up, demonstrating durable and sustainable solutions. Experience building self-healing or auto-remediation platforms (remote actions, osquery + response, custom agents). Exposure to OT (Operational Technology) systems. Comfort operating in an SRE culture: SLOs, error budgets, and blameless postmortems applied to the end-user experience. Compensation For this role, the target salary range is $164,000 - $270,000 (actual range may vary based on experience). This is the lowest to highest salary we reasonably and in good faith believe we would pay for this role at the time of this posting. We may ultimately pay more or less than the posted range, and the range may be modified in the future. An employee's pay position within the salary range will be based on several factors, including, but not limited to, relevant education, qualifications, certifications, experience, skills, geographic location, performance, and business or organizational needs. Benefits for Full-time Employees Medical, dental, vision, and life insurance plans for employees 401k Relocation support may be provided for certain situations, based on business need. Flexible vacation policy Equity ITAR Requirements To conform to U.S. Government space technology export regulations, including the International Traffic in Arms Regulations (ITAR) you must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here. Hadrian Is An Equal Opportunity Employer It is the Company's policy to provide equal employment opportunity for all applicants and employees. The Company does not unlawfully discriminate on the basis of race inclusive of traits historically associated with race (including, but not limited to, hair texture and protective hairstyles, such as braids, locks and twists), color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender identity, gender expression, transgender status, national origin (including, in California, possession of a drivers license), ancestry, citizenship, age, physical or mental disability, height or weight, medical condition, family care status, military or veteran status, marital status, domestic partner status, sexual orientation, genetic information, exercise of reproductive rights, any other basis protected by local, state, or federal laws, or any combination of the above characteristics. When necessary, the Company also makes reasonable accommodations for disabled candidates and employees, including for candidates or employees who are disabled by pregnancy, childbirth, or related medical conditions.
Sr. Site Reliability Engineer
PayNearMe, Inc. Alviso, California
Job Description Job Description Company Description At PayNearMe, we're on a mission to make paying and getting paid as simple as possible. We build innovative technology that transforms the way businesses and their customers experience payments. Our industry-leading platform, PayXM , is the first of its kind-designed to manage the entire payment experience from start to finish. Every click, swipe or tap is seamless, fast and secure, helping non-commerce businesses boost customer satisfaction, accelerate payments, and reduce costs. Our single platform handles it all: cards, ACH, digital wallets such as PayPal, Venmo, Cash App Pay, Apple Pay and Google Pay, and even cash at more than 62,000 retail locations nationwide. Today, thousands of businesses across consumer lending, iGaming and online sports betting, property management, and tolling trust PayNearMe to deliver a payment experience that drives real results. In September 2025, we raised a $50 million Series E funding round to accelerate our growth. We're a team of 300+ employees across 41 states, headquartered in Silicon Valley with satellite offices in Dallas, TX and Holmdel, NJ. Join us and be part of a team that's shaping the future of payments-one experience at a time. As our Site Reliability Engineer, you will design, build, and maintain the systems and infrastructure that power our applications, ensuring their reliability, scalability, and performance. You will bring a software engineering approach to operations, automating processes, and continuously improving the infrastructure and tools to support our business needs. Responsibilities Infrastructure Management: Design, implement, and maintain scalable and resilient infrastructure using Terraform for infrastructure as code, ensuring high availability and performance Kubernetes and Containers: Deploy, manage, and optimize Kubernetes clusters and containerized applications using Docker. Implement best practices for container orchestration and management Systems and Application Monitoring/Observability: Develop and maintain comprehensive monitoring and observability solutions using Datadog. Ensure detailed visibility into system performance and application health SLOs and SLA Management: Define, monitor, and maintain Service Level Objectives (SLOs) and Service Level Agreements (SLAs) to ensure reliable and consistent service delivery Incident Response and Troubleshooting: Respond to incidents, perform root cause analysis, and implement solutions to prevent recurrence. Participate in post-incident reviews and contribute to blameless postmortems Reliability and Production Environment Management: Ensure the reliability and stability of our production environments. Continuously assess and improve system reliability, identifying and addressing potential points of failure Automation and Scripting: Develop automation scripts and tools to reduce manual intervention and improve system reliability using Python, Bash, or Go. Implement and improve CI/CD pipelines CI/CD Pipeline Management: Enhance and maintain continuous integration and continuous deployment pipelines using GitLab CI. Ensure seamless and reliable deployment processes Capacity Planning and Scaling: Assist in capacity planning and ensure that systems are scalable to meet future demands. Implement auto-scaling strategies where applicable Security and Compliance: Implement security best practices and ensure compliance with industry standards. Regularly review and update security policies and procedures Collaboration and Support: Work closely with development teams to ensure reliability and scalability of new features and services. Provide technical support and guidance on infrastructure-related issues Software Engineering for Operations: Develop and maintain internal tools and services that enhance the efficiency and reliability of our operations On-Call Rotation: Participate in an on-call rotation to address production issues and collaborate in incident response efforts Qualifications +3 years of experience in SRE, DevOps, or a related role Cloud Platform Experience: Proficient with cloud platforms such as AWS, GCP, or Azure Experience with EC2, RDS, VPCs, and security groups is essential. Kubernetes and Containers: Strong experience with Kubernetes and Docker, including deployment, scaling, and management of containerized applications Infrastructure as Code: Expert in using Terraform for infrastructure as code. Proficient with configuration management tools such as Ansible, Puppet, or Chef Monitoring and Observability: Extensive experience with monitoring and observability tools like Datadog, Prometheus, Grafana, ELK stack, or Splunk. Skilled in setting up detailed monitoring and logging systems SLOs and SLA Management: Proven ability to define, monitor, and maintain SLOs and SLAs to ensure reliable service delivery Scripting and Automation: Strong skills in scripting languages like Python, Bash, or Go. Experience automating repetitive tasks and processes CI/CD Practices: Familiarity with GitLab CI or similar tool for continuous integration and deployment. Experience in setting up and managing pipelines Production Environments: Experience supporting production environments running Go or Ruby/Rails applications Tool Development: Ability to write and update tools to support infrastructure and application management, demonstrating the principle that "SRE is what happens when you ask a software engineer to design an operations team DevOps Best Practices: Deep understanding of DevOps principles, practices, and tools to drive continuous improvement in the software development lifecycle Soft Skills: Strong organizational skills, attention to detail, and the ability to work collaboratively in a team environment. Excellent documentation skills to ensure accurate and detailed records Problem-Solving Ability: Excellent analytical and problem-solving skills to diagnose and resolve complex system issues quickly and effectively The annual base salary range for this role represents PayNearMe's good-faith estimate of the base salary it reasonably expects to offer for this position at the time of hire. Actual compensation may vary based on factors including the candidate's experience, qualifications, skills, and work location. PayNearMe may offer compensation outside of this range in certain circumstances. This position will remain posted until filled. Annual Salary Range $180,000-$200,000 USD Why Join Us?: Competitive salary and benefits with growth-company options grant Fast- paced and professional work culture Stock options with standard startup vesting - 1 year cliff; 4 years total $50 monthly communication expense stipend to go towards your phone/internet bill $250 stipend to enhance your WFH setup Reimbursement for peripheral equipment: monitor (up to $400), keyboard and mouse (up to $200) Premium medical benefits including vision and dental (100% coverage for employees) Company-sponsored life and disability insurance Paid parental bonding leave Paid sick leave, jury duty, bereavement 401k plan Flexible Time Off (our team members typically take off 3-4 weeks per year) Volunteer Time Off 13 scheduled holidays PayNearMe strives to create a workplace where all employees thrive. Our core values represent who we are today and we take pride in the way we work with each other as well as with our stakeholders. We're in this together to do the right thing. We deliver real results we are proud of while remaining respectful, transparent, and flexible. PayNearMe is an equal opportunity employer. We are diligently and thoughtfully working towards cultivating a diverse workforce which in turn, enhances our products and services for the communities we serve. Applicants who represent all backgrounds are strongly encouraged to apply. CALIFORNIA CONSUMER PRIVACY ACT: APPLICANT NOTICE Effective Date: January 1, 2020 Last Reviewed on: December 23, 2019 PayNearMe, Inc. (the "Company") is providing you with this Notice ("Notice") to inform you about: the categories of Personal Information that the Company collects and maintains about applicants; and the purposes for which the Company uses that Personal Information. For purposes of this Notice, "Personal Information" means information that identifies, relates to, describes, is capable of being associated with, or could reasonably be linked, directly or indirectly with, a natural person that the Company may collect in connection with screening applicants for job openings at the Company. Identifiers and Professional or Employment-Related Information. The Company collects identifiers and professional or employment-related information, which may include some or all the following: real name, nickname or alias, postal address, telephone number, e-mail address, membership in professional organizations, professional certifications, language skills, and current and past employment history. The Company collects this Personal Information to evaluate previous job performance and consider applicants for positions, to develop a talent pool and plan for succession . click apply for full job details
08/05/2026
Full time
Job Description Job Description Company Description At PayNearMe, we're on a mission to make paying and getting paid as simple as possible. We build innovative technology that transforms the way businesses and their customers experience payments. Our industry-leading platform, PayXM , is the first of its kind-designed to manage the entire payment experience from start to finish. Every click, swipe or tap is seamless, fast and secure, helping non-commerce businesses boost customer satisfaction, accelerate payments, and reduce costs. Our single platform handles it all: cards, ACH, digital wallets such as PayPal, Venmo, Cash App Pay, Apple Pay and Google Pay, and even cash at more than 62,000 retail locations nationwide. Today, thousands of businesses across consumer lending, iGaming and online sports betting, property management, and tolling trust PayNearMe to deliver a payment experience that drives real results. In September 2025, we raised a $50 million Series E funding round to accelerate our growth. We're a team of 300+ employees across 41 states, headquartered in Silicon Valley with satellite offices in Dallas, TX and Holmdel, NJ. Join us and be part of a team that's shaping the future of payments-one experience at a time. As our Site Reliability Engineer, you will design, build, and maintain the systems and infrastructure that power our applications, ensuring their reliability, scalability, and performance. You will bring a software engineering approach to operations, automating processes, and continuously improving the infrastructure and tools to support our business needs. Responsibilities Infrastructure Management: Design, implement, and maintain scalable and resilient infrastructure using Terraform for infrastructure as code, ensuring high availability and performance Kubernetes and Containers: Deploy, manage, and optimize Kubernetes clusters and containerized applications using Docker. Implement best practices for container orchestration and management Systems and Application Monitoring/Observability: Develop and maintain comprehensive monitoring and observability solutions using Datadog. Ensure detailed visibility into system performance and application health SLOs and SLA Management: Define, monitor, and maintain Service Level Objectives (SLOs) and Service Level Agreements (SLAs) to ensure reliable and consistent service delivery Incident Response and Troubleshooting: Respond to incidents, perform root cause analysis, and implement solutions to prevent recurrence. Participate in post-incident reviews and contribute to blameless postmortems Reliability and Production Environment Management: Ensure the reliability and stability of our production environments. Continuously assess and improve system reliability, identifying and addressing potential points of failure Automation and Scripting: Develop automation scripts and tools to reduce manual intervention and improve system reliability using Python, Bash, or Go. Implement and improve CI/CD pipelines CI/CD Pipeline Management: Enhance and maintain continuous integration and continuous deployment pipelines using GitLab CI. Ensure seamless and reliable deployment processes Capacity Planning and Scaling: Assist in capacity planning and ensure that systems are scalable to meet future demands. Implement auto-scaling strategies where applicable Security and Compliance: Implement security best practices and ensure compliance with industry standards. Regularly review and update security policies and procedures Collaboration and Support: Work closely with development teams to ensure reliability and scalability of new features and services. Provide technical support and guidance on infrastructure-related issues Software Engineering for Operations: Develop and maintain internal tools and services that enhance the efficiency and reliability of our operations On-Call Rotation: Participate in an on-call rotation to address production issues and collaborate in incident response efforts Qualifications +3 years of experience in SRE, DevOps, or a related role Cloud Platform Experience: Proficient with cloud platforms such as AWS, GCP, or Azure Experience with EC2, RDS, VPCs, and security groups is essential. Kubernetes and Containers: Strong experience with Kubernetes and Docker, including deployment, scaling, and management of containerized applications Infrastructure as Code: Expert in using Terraform for infrastructure as code. Proficient with configuration management tools such as Ansible, Puppet, or Chef Monitoring and Observability: Extensive experience with monitoring and observability tools like Datadog, Prometheus, Grafana, ELK stack, or Splunk. Skilled in setting up detailed monitoring and logging systems SLOs and SLA Management: Proven ability to define, monitor, and maintain SLOs and SLAs to ensure reliable service delivery Scripting and Automation: Strong skills in scripting languages like Python, Bash, or Go. Experience automating repetitive tasks and processes CI/CD Practices: Familiarity with GitLab CI or similar tool for continuous integration and deployment. Experience in setting up and managing pipelines Production Environments: Experience supporting production environments running Go or Ruby/Rails applications Tool Development: Ability to write and update tools to support infrastructure and application management, demonstrating the principle that "SRE is what happens when you ask a software engineer to design an operations team DevOps Best Practices: Deep understanding of DevOps principles, practices, and tools to drive continuous improvement in the software development lifecycle Soft Skills: Strong organizational skills, attention to detail, and the ability to work collaboratively in a team environment. Excellent documentation skills to ensure accurate and detailed records Problem-Solving Ability: Excellent analytical and problem-solving skills to diagnose and resolve complex system issues quickly and effectively The annual base salary range for this role represents PayNearMe's good-faith estimate of the base salary it reasonably expects to offer for this position at the time of hire. Actual compensation may vary based on factors including the candidate's experience, qualifications, skills, and work location. PayNearMe may offer compensation outside of this range in certain circumstances. This position will remain posted until filled. Annual Salary Range $180,000-$200,000 USD Why Join Us?: Competitive salary and benefits with growth-company options grant Fast- paced and professional work culture Stock options with standard startup vesting - 1 year cliff; 4 years total $50 monthly communication expense stipend to go towards your phone/internet bill $250 stipend to enhance your WFH setup Reimbursement for peripheral equipment: monitor (up to $400), keyboard and mouse (up to $200) Premium medical benefits including vision and dental (100% coverage for employees) Company-sponsored life and disability insurance Paid parental bonding leave Paid sick leave, jury duty, bereavement 401k plan Flexible Time Off (our team members typically take off 3-4 weeks per year) Volunteer Time Off 13 scheduled holidays PayNearMe strives to create a workplace where all employees thrive. Our core values represent who we are today and we take pride in the way we work with each other as well as with our stakeholders. We're in this together to do the right thing. We deliver real results we are proud of while remaining respectful, transparent, and flexible. PayNearMe is an equal opportunity employer. We are diligently and thoughtfully working towards cultivating a diverse workforce which in turn, enhances our products and services for the communities we serve. Applicants who represent all backgrounds are strongly encouraged to apply. CALIFORNIA CONSUMER PRIVACY ACT: APPLICANT NOTICE Effective Date: January 1, 2020 Last Reviewed on: December 23, 2019 PayNearMe, Inc. (the "Company") is providing you with this Notice ("Notice") to inform you about: the categories of Personal Information that the Company collects and maintains about applicants; and the purposes for which the Company uses that Personal Information. For purposes of this Notice, "Personal Information" means information that identifies, relates to, describes, is capable of being associated with, or could reasonably be linked, directly or indirectly with, a natural person that the Company may collect in connection with screening applicants for job openings at the Company. Identifiers and Professional or Employment-Related Information. The Company collects identifiers and professional or employment-related information, which may include some or all the following: real name, nickname or alias, postal address, telephone number, e-mail address, membership in professional organizations, professional certifications, language skills, and current and past employment history. The Company collects this Personal Information to evaluate previous job performance and consider applicants for positions, to develop a talent pool and plan for succession . click apply for full job details
Sr. Lead, Machine Learning Engineer (Enterprise Platforms Technology)
Capital One Mc Lean, Virginia
Sr. Lead, Machine Learning Engineer (Enterprise Platforms Technology) As a Capital One Machine Learning Engineer (MLE), you'll be part of an Agile team dedicated to productionizing machine learning applications and systems at scale. You'll participate in the detailed technical design, development, and implementation of machine learning applications using existing and emerging technology platforms. You'll focus on machine learning architectural design, develop and review model and application code, and ensure high availability and performance of our machine learning applications. You'll have the opportunity to continuously learn and apply the latest innovations and best practices in machine learning engineering. Enterprise Platforms Technology (EPTech) comprises many of Capital One's most important enterprise platforms. We play an essential role in establishing practices for building technology solutions across the company, while also delivering capabilities that exemplify those practices. What you'll do in the role: The MLE role overlaps with many disciplines, such as Ops, Modeling, and Data Engineering. In this role, you'll be expected to perform many ML engineering activities, including one or more of the following: Design, build, and/or deliver ML models and components that solve real-world business problems, while working in collaboration with the Product and Data Science teams. Inform your ML infrastructure decisions using your understanding of ML modeling techniques and issues, including choice of model, data, and feature selection, model training, hyperparameter tuning, dimensionality, bias/variance, and validation). Solve complex problems by writing and testing application code, developing and validating ML models, and automating tests and deployment. Collaborate as part of a cross-functional Agile team to create and enhance software that enables state-of-the-art big data and ML applications. Retrain, maintain, and monitor models in production. Leverage or build cloud-based architectures, technologies, and/or platforms to deliver optimized ML models at scale. Construct optimized data pipelines to feed ML models. Leverage continuous integration and continuous deployment best practices, including test automation and monitoring, to ensure successful deployment of ML models and application code. Ensure all code is well-managed to reduce vulnerabilities, models are well-governed from a risk perspective, and the ML follows best practices in Responsible and Explainable AI. Use programming languages like Python, Scala, or Java. Basic Qualifications: Bachelor's degree At least 8 years of experience designing and building data-intensive solutions using distributed computing (Internship experience does not apply) At least 4 years of experience programming with Python, Scala, or Java At least 3 years of experience building, scaling, and optimizing ML systems At least 2 years of experience leading teams developing ML solutions Preferred Qualifications: Master's or doctoral degree in computer science, electrical engineering, mathematics, or a similar field Experience developing and deploying ML solutions in a public cloud such as AWS, Azure, or Google Cloud Platform 4+ years of on-the-job experience with an industry recognized ML framework such as scikit-learn, PyTorch, Dask, Spark, or TensorFlow 3+ years of experience developing performant, resilient, and maintainable code 3+ years of experience with data gathering and preparation for ML models 3+ years of people management experience ML industry impact through conference presentations, papers, blog posts, open source contributions, or patents 3+ years of experience building production-ready data pipelines that feed ML models Ability to communicate complex technical concepts clearly to a variety of audiences Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $229,900 - $262,400 for Sr. Lead Machine Learning Engineer New York, NY: $250,800 - $286,200 for Sr. Lead Machine Learning Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
08/05/2026
Full time
Sr. Lead, Machine Learning Engineer (Enterprise Platforms Technology) As a Capital One Machine Learning Engineer (MLE), you'll be part of an Agile team dedicated to productionizing machine learning applications and systems at scale. You'll participate in the detailed technical design, development, and implementation of machine learning applications using existing and emerging technology platforms. You'll focus on machine learning architectural design, develop and review model and application code, and ensure high availability and performance of our machine learning applications. You'll have the opportunity to continuously learn and apply the latest innovations and best practices in machine learning engineering. Enterprise Platforms Technology (EPTech) comprises many of Capital One's most important enterprise platforms. We play an essential role in establishing practices for building technology solutions across the company, while also delivering capabilities that exemplify those practices. What you'll do in the role: The MLE role overlaps with many disciplines, such as Ops, Modeling, and Data Engineering. In this role, you'll be expected to perform many ML engineering activities, including one or more of the following: Design, build, and/or deliver ML models and components that solve real-world business problems, while working in collaboration with the Product and Data Science teams. Inform your ML infrastructure decisions using your understanding of ML modeling techniques and issues, including choice of model, data, and feature selection, model training, hyperparameter tuning, dimensionality, bias/variance, and validation). Solve complex problems by writing and testing application code, developing and validating ML models, and automating tests and deployment. Collaborate as part of a cross-functional Agile team to create and enhance software that enables state-of-the-art big data and ML applications. Retrain, maintain, and monitor models in production. Leverage or build cloud-based architectures, technologies, and/or platforms to deliver optimized ML models at scale. Construct optimized data pipelines to feed ML models. Leverage continuous integration and continuous deployment best practices, including test automation and monitoring, to ensure successful deployment of ML models and application code. Ensure all code is well-managed to reduce vulnerabilities, models are well-governed from a risk perspective, and the ML follows best practices in Responsible and Explainable AI. Use programming languages like Python, Scala, or Java. Basic Qualifications: Bachelor's degree At least 8 years of experience designing and building data-intensive solutions using distributed computing (Internship experience does not apply) At least 4 years of experience programming with Python, Scala, or Java At least 3 years of experience building, scaling, and optimizing ML systems At least 2 years of experience leading teams developing ML solutions Preferred Qualifications: Master's or doctoral degree in computer science, electrical engineering, mathematics, or a similar field Experience developing and deploying ML solutions in a public cloud such as AWS, Azure, or Google Cloud Platform 4+ years of on-the-job experience with an industry recognized ML framework such as scikit-learn, PyTorch, Dask, Spark, or TensorFlow 3+ years of experience developing performant, resilient, and maintainable code 3+ years of experience with data gathering and preparation for ML models 3+ years of people management experience ML industry impact through conference presentations, papers, blog posts, open source contributions, or patents 3+ years of experience building production-ready data pipelines that feed ML models Ability to communicate complex technical concepts clearly to a variety of audiences Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. McLean, VA: $229,900 - $262,400 for Sr. Lead Machine Learning Engineer New York, NY: $250,800 - $286,200 for Sr. Lead Machine Learning Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
Senior Lead AI Engineer (MLXT)
Capital One New York, New York
Senior Lead AI Engineer (MLXT) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. As a Senior Lead AI Engineer, you will drive critical technical initiatives that shape the future of enterprise AIML infrastructure targeted to transform the business analysis experience in a highly regulated environment. You will play a pivotal role in modernizing our core Analyst & AIML workflow orchestration layer and user interface, integrating frontier generative AI capabilities and self-serve tooling to enable no-code/low-code, well-governed model development for business domain experts. In this role, you will: Modernize Workflows: Advance the orchestration layer and UI with generative AI capabilities and self-serve tooling to democratize AI development through governed, low-code/no-code tools for business domain experts. Build Agentic-Driven Infrastructure: Modernize our platform services with agentic infrastructure that is reliable, scalable, secure, and seamlessly adheres to enterprise guardrails. Drive AutoML Innovation: Scale enterprise-grade managed AutoML offerings for tabular and time-series data to radically reduce solution time-to-market from weeks to days. Engineer AI-Driven Controls: Evolve the core component marketplace by engineering cutting-edge, automated governance frameworks The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good. Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production. You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven. You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enable you to see and exploit optimization opportunities that others miss. You are a resilient trail blazer who can forge new paths to achieve business goals when the route is unknown. Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, or Java Preferred Qualifications: 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, integrating, delivering, and supporting complex AI systems Demonstrated ability to lead and mentor an engineering team and influence cross-functional stakeholders Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for Sr. Lead AI Engineer McLean, VA: $229,900 - $262,400 for Sr. Lead AI Engineer New York, NY: $250,800 - $286,200 for Sr. Lead AI Engineer San Francisco, CA: $250,800 - $286,200 for Sr. Lead AI Engineer San Jose, CA: $250,800 - $286,200 for Sr. Lead AI Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
08/05/2026
Full time
Senior Lead AI Engineer (MLXT) Overview: At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Intelligent Foundations and Experiences (IFX) team is at the center of bringing our vision for AI at Capital One to life. We work hand-in-hand with our partners across the company to advance the state of the art in science and AI engineering, and we build and deploy proprietary solutions that are central to our business and deliver value to millions of customers. Our AI models and platforms empower teams across Capital One to enhance their products with the transformative power of AI, in responsible and scalable ways for the highest leverage impact. As a Senior Lead AI Engineer, you will drive critical technical initiatives that shape the future of enterprise AIML infrastructure targeted to transform the business analysis experience in a highly regulated environment. You will play a pivotal role in modernizing our core Analyst & AIML workflow orchestration layer and user interface, integrating frontier generative AI capabilities and self-serve tooling to enable no-code/low-code, well-governed model development for business domain experts. In this role, you will: Modernize Workflows: Advance the orchestration layer and UI with generative AI capabilities and self-serve tooling to democratize AI development through governed, low-code/no-code tools for business domain experts. Build Agentic-Driven Infrastructure: Modernize our platform services with agentic infrastructure that is reliable, scalable, secure, and seamlessly adheres to enterprise guardrails. Drive AutoML Innovation: Scale enterprise-grade managed AutoML offerings for tabular and time-series data to radically reduce solution time-to-market from weeks to days. Engineer AI-Driven Controls: Evolve the core component marketplace by engineering cutting-edge, automated governance frameworks The Ideal Candidate: You love to build systems, take pride in the quality of your work, and also share our passion to do the right thing. You want to work on problems that will help change banking for good. Passion for staying abreast of the latest research, and an ability to intuitively understand scientific publications and judiciously apply novel techniques in production. You adapt quickly and thrive on bringing clarity to big, undefined problems. You love asking questions and digging deep to uncover the root of problems and can articulate your findings concisely with clarity. You have the courage to share new ideas even when they are unproven. You are deeply Technical. You possess a strong foundation in engineering and mathematics, and your expertise in hardware, software, and AI enable you to see and exploit optimization opportunities that others miss. You are a resilient trail blazer who can forge new paths to achieve business goals when the route is unknown. Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 6 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies At least 6 years of experience programming with Python, Go, Scala, or Java Preferred Qualifications: 7 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g. AWS, Google Cloud, Azure, or equivalent private cloud) Experience designing, developing, integrating, delivering, and supporting complex AI systems Demonstrated ability to lead and mentor an engineering team and influence cross-functional stakeholders Experience developing AI and ML algorithms or technologies (e.g. LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost Passion for staying abreast of the latest AI research and AI systems, and judiciously apply novel techniques in production Excellent communication and presentation skills, with the ability to articulate complex AI concepts to peers Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Cambridge, MA: $229,900 - $262,400 for Sr. Lead AI Engineer McLean, VA: $229,900 - $262,400 for Sr. Lead AI Engineer New York, NY: $250,800 - $286,200 for Sr. Lead AI Engineer San Francisco, CA: $250,800 - $286,200 for Sr. Lead AI Engineer San Jose, CA: $250,800 - $286,200 for Sr. Lead AI Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
Senior Distinguished Engineer, AI Compute (Remote Eligible)
Capital One Richmond, Virginia
Senior Distinguished Engineer, AI Compute (Remote Eligible) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Capital One machine learning platform organization manages our cloud-based enterprise AI+ML system delivering the high-scale developer and runtime environments required to build, orchestrate, and deploy compute and data intensive AI systems across real-time and batch workloads. We are seeking a Senior Distinguished Engineer, a hands-on technical leader passionate about distributed systems, to engineer and scale foundational compute capabilities for our platform. You will use your experience in building large scale, highly available and high performance systems to develop our common compute infrastructure on top of CPU and GPU substrates. Your contributions will power everything from developer notebooks to ML / DL model training, model inference and feature generation pipelines to pre-training and fine tuning Transformer-based models as well as generative AI inference and agentic applications. Your depth of expertise in technologies including Golang and Python programming languages, popular distributed compute frameworks including Spark / Dask / Ray / Flink, container (e.g., Kubernetes) and serverless (e.g., AWS Lambda) runtime environments, and ML+AI workload patterns will provide an amplifying technical element that is paramount to our team's success. In this role, you will : Architect and build control and data plane implementations required to realize a highly available, multi-tenant, large scale and a secure machine learning platform Develop Ray and Spark distributed compute engine solutions to accelerate diverse workloads from LLM pre-training and reinforcement learning to large-scale data processing, while maximizing compute unit economics Engineer systemic improvements for operational excellence including automating KTLO (Keep The Lights On) workflows Direct the technical execution of a diverse project portfolio, collaborating with developers specializing in everything ranging from distributed microservices to running large foundation models Work cross-functionally with product and program management disciplines, and stakeholder and partners across Capital One to help optimize business outcomes while driving towards strong technology solutions Share your passion for staying on top of tech trends, experimenting with and learning new technologies, participating in internal & external technology communities, and leading system design and code review sessions Help elevate the Capital One Distinguished Engineering community and establish yourself as a go-to resource on given technologies and technology-enabled capabilities Lead the way in creating next-generation talent, mentoring internal talent and actively recruiting external talent to bolster the Capital One tech talent pool Capital One is open to hiring a Remote Employee for this opportunity Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 10 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 8 years of experience developing AI and ML algorithms or technologies At least 10 years of experience programming with Python, Go, Scala, or Java Preferred Qualifications : Master's Degree in Computer Science or a Master's Degree in Software Engineering Hands on experience in the internals of Ray (Actors/GCS/Scheduling) or Spark (Query Optimizer/Memory Management) Experience building platforms that support LLM training, fine-tuning, or high-throughput inference Hands-on experience with AWS-specific compute primitives (EKS, EC2 UltraClusters, Graviton) and cost-optimization strategies History of upstream contributions to major distributed systems projects Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Remote (Regardless of Location): $286,200 - $326,700 for Sr. Distinguished AI Engineer Cambridge, MA: $314,800 - $359,300 for Sr. Distinguished AI Engineer McLean, VA: $314,800 - $359,300 for Sr. Distinguished AI Engineer New York, NY: $343,400 - $392,000 for Sr. Distinguished AI Engineer Richmond, VA: $286,200 - $326,700 for Sr. Distinguished AI Engineer San Francisco, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer San Jose, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
08/05/2026
Full time
Senior Distinguished Engineer, AI Compute (Remote Eligible) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized customer experiences. Our investments in technology infrastructure and world-class talent - along with our deep experience in machine learning - position us to be at the forefront of enterprises leveraging AI. From informing customers about unusual charges to answering their questions in real time, our applications of AI & ML are bringing humanity and simplicity to banking. We are committed to continuing to build world-class applied science and engineering teams to deliver our industry leading capabilities with breakthrough product experiences and scalable, high-performance AI infrastructure. At Capital One, you will help bring the transformative power of emerging AI capabilities to reimagine how we serve our customers and businesses who have come to love the products and services we build. Team Description: The Capital One machine learning platform organization manages our cloud-based enterprise AI+ML system delivering the high-scale developer and runtime environments required to build, orchestrate, and deploy compute and data intensive AI systems across real-time and batch workloads. We are seeking a Senior Distinguished Engineer, a hands-on technical leader passionate about distributed systems, to engineer and scale foundational compute capabilities for our platform. You will use your experience in building large scale, highly available and high performance systems to develop our common compute infrastructure on top of CPU and GPU substrates. Your contributions will power everything from developer notebooks to ML / DL model training, model inference and feature generation pipelines to pre-training and fine tuning Transformer-based models as well as generative AI inference and agentic applications. Your depth of expertise in technologies including Golang and Python programming languages, popular distributed compute frameworks including Spark / Dask / Ray / Flink, container (e.g., Kubernetes) and serverless (e.g., AWS Lambda) runtime environments, and ML+AI workload patterns will provide an amplifying technical element that is paramount to our team's success. In this role, you will : Architect and build control and data plane implementations required to realize a highly available, multi-tenant, large scale and a secure machine learning platform Develop Ray and Spark distributed compute engine solutions to accelerate diverse workloads from LLM pre-training and reinforcement learning to large-scale data processing, while maximizing compute unit economics Engineer systemic improvements for operational excellence including automating KTLO (Keep The Lights On) workflows Direct the technical execution of a diverse project portfolio, collaborating with developers specializing in everything ranging from distributed microservices to running large foundation models Work cross-functionally with product and program management disciplines, and stakeholder and partners across Capital One to help optimize business outcomes while driving towards strong technology solutions Share your passion for staying on top of tech trends, experimenting with and learning new technologies, participating in internal & external technology communities, and leading system design and code review sessions Help elevate the Capital One Distinguished Engineering community and establish yourself as a go-to resource on given technologies and technology-enabled capabilities Lead the way in creating next-generation talent, mentoring internal talent and actively recruiting external talent to bolster the Capital One tech talent pool Capital One is open to hiring a Remote Employee for this opportunity Basic Qualifications: Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 10 years of experience developing AI and ML algorithms or technologies, or a Master's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 8 years of experience developing AI and ML algorithms or technologies At least 10 years of experience programming with Python, Go, Scala, or Java Preferred Qualifications : Master's Degree in Computer Science or a Master's Degree in Software Engineering Hands on experience in the internals of Ray (Actors/GCS/Scheduling) or Spark (Query Optimizer/Memory Management) Experience building platforms that support LLM training, fine-tuning, or high-throughput inference Hands-on experience with AWS-specific compute primitives (EKS, EC2 UltraClusters, Graviton) and cost-optimization strategies History of upstream contributions to major distributed systems projects Capital One will consider sponsoring a new qualified applicant for employment authorization for this position. The minimum and maximum full-time annual salaries for this role are listed below, by location. Please note that this salary information is solely for candidates hired to perform work within one of these locations, and refers to the amount Capital One is willing to pay at the time of this posting. Salaries for part-time roles will be prorated based upon the agreed upon number of hours to be regularly worked. Remote (Regardless of Location): $286,200 - $326,700 for Sr. Distinguished AI Engineer Cambridge, MA: $314,800 - $359,300 for Sr. Distinguished AI Engineer McLean, VA: $314,800 - $359,300 for Sr. Distinguished AI Engineer New York, NY: $343,400 - $392,000 for Sr. Distinguished AI Engineer Richmond, VA: $286,200 - $326,700 for Sr. Distinguished AI Engineer San Francisco, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer San Jose, CA: $343,400 - $392,000 for Sr. Distinguished AI Engineer Candidates hired to work in other locations will be subject to the pay range associated with that location, and the actual annualized salary amount offered to any candidate at the time of hire will be reflected solely in the candidate's offer letter. This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI). Incentives could be discretionary or non discretionary depending on the plan. Capital One offers a comprehensive, competitive, and inclusive set of health, financial and other benefits that support your total well-being. Learn more at the Capital One Careers website . Eligibility varies based on full or part-time status, exempt or non-exempt status, and management level. This role is expected to accept applications for a minimum of 5 business days.No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace. Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with the requirements of applicable laws regarding criminal background inquiries, including, to the extent applicable, Article 23-A of the New York Correction Law; San Francisco, California Police Code Article 49, Sections ; New York City's Fair Chance Act; Philadelphia's Fair Criminal Records Screening Act; and other applicable federal, state, and local laws and regulations regarding criminal background inquiries. If you have visited our website in search of information on employment opportunities or to apply for a position, and you require an accommodation, please contact Capital One Recruiting at 1- or via email at . All information you provide will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations. For technical support or questions about Capital One's recruiting process, please send an email to Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site. Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe and any position posted in the Philippines is for Capital One Philippines Service Corp. (COPSSC).
Senior Site Reliability Engineer- San Francisco, CA, the US
Kody San Francisco, California
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
08/05/2026
Full time
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board