it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

5 jobs found

Email me jobs like this
Refine Search
Current Search
devops engineer experienced or senior
Technicial Product Owner
Brinks Home
Description Brinks Home is a leader in the smart security industry, protecting over one million people across the U.S., Canada, and Puerto Rico. Our platinum-grade protection is backed by award-winning customer service and expertly trained professionals. We strive for the highest standards for our customers while fostering a positive work environment for our employees. We create a culture that fosters innovation, celebrates creativity, and encourages authenticity. Join us and be part of a collaborative team that is relentless in our pursuit of security for life. Position Overview: We are currently seeking a seasoned, results-driven Technical Product Owner who embodies our core values: Service, Accountability, Customer Focus, Growth, and Integrity. This role will lead the strategy, development, and optimization of our sales and field service experiences across mobile and web - the tools our field technicians rely on to install, activate, and service equipment in customers' homes every day. You will ensure our products deliver exceptional value to users, with a focus on usability, performance, and innovation. You will operate as the Product Owner on a Scrum team within a SAFe environment, working shoulder-to-shoulder with Field Operations, Engineering, and Sales, and using AI to work faster and smarter across the entire product lifecycle. As a senior individual contributor, you will set direction in ambiguous problem spaces with minimal oversight, influence without authority, and raise the bar for product practice across the team. Key Responsibilities: Own and manage the mobile and web product backlog, equipment inventory capabilities, and scheduling and dispatch capabilities. Prioritize features that enhance the user experience across both platforms. Define clear user stories, acceptance criteria, and success metrics with instrumentation defined up front, then use experimentation - A/B tests, feature flags, and phased rollouts - and post-release data to measure adoption and iterate. Collaborate with UX/UI designers to ensure intuitive, responsive, and accessible mobile and web interfaces. Advocate for mobile-first, responsive design principles across the organization and stay current on mobile and web trends, technologies, and best practices. Partner with Field Operations leadership and field technicians to translate equipment job installation, activation, and service workflows into clear product requirements. Own and communicate the product roadmap for your area, sequencing releases against business outcomes, keeping stakeholders current on status and expectations, and making trade-offs visible to leadership. Identify, pilot, and champion AI tools that accelerate discovery, backlog refinement, documentation, and day-to-day team productivity. Define requirements for AI-enabled product features such as guided troubleshooting, intelligent scheduling, and automated job verification - including data quality, human-in-the-loop review, and responsible-AI guardrails. Serve as Product Owner on a Scrum team operating within a SAFe cadence - PI planning, backlog refinement, iteration planning, reviews, and system demos. Manage cross-team dependencies and integration requirements across CRM, dispatch and scheduling, billing, e-signature, order management and inventory, and device telemetry systems. Partner with developers across the full software development lifecycle, from discovery through release - technical feasibility, timely delivery, QA, UAT, and deployment readiness. Define non-functional requirements with engineering - performance, offline and low-connectivity resilience, security, accessibility, and data-privacy compliance. Drive alignment across Field Ops, Sales, Engineering, QA, Training, and Customer Care in a fast-paced, cross-functional environment. Own release readiness and adoption: release notes, field enablement and training content, pilot and rollout plans, and post-launch issue triage with engineering. Required Qualifications: 7+ years of experience as a Product Owner, with a strong focus on mobile and web applications, including at least 3 years owning a technical product alongside an engineering team. 5+ years of experience with SAFe Agile methodology and Jira. AI-savvy - strong understanding of AI/ML for identifying opportunities, requirements, and design aligned with business goals, plus practical use of AI in day-to-day work to maximize individual and team productivity. Hands-on experience as a Product Owner on a Scrum team, including representing a product area across multiple PI planning cycles; SAFe (Scaled Agile Framework) experience strongly preferred. Command of the software development lifecycle and its supporting functions, including QA, release management, DevOps, documentation, and training. Strong requirements craft - user story mapping, process and workflow mapping, clear written specifications, and testable acceptance criteria. Proven track record of delivering successful, high-usage mobile (iOS and/or Android) and web applications from concept through multiple release cycles. Data fluency - comfortable querying data or working in BI dashboards, and using funnel, cohort, and adoption analysis to size opportunities and defend decisions. Ability to operate autonomously - setting direction in ambiguity, making defensible trade-off decisions, and driving outcomes end to end. Excellent communication, stakeholder management, and problem-solving skills, with the ability to communicate effectively across all levels of operations and users. Proven ability to thrive on fast-paced, cross-functional teams with competing priorities and shifting requirements. Preferred Qualifications: Experience within the home security industry and/or a field operations dispatch. Working knowledge of scheduling capabilities and inventory management. Experience with A/B testing and product analytics across mobile and web. Proficiency with backlog management, documentation, and design collaboration tooling. Experience mentoring or coaching less-experienced product owners and analysts, and influencing senior stakeholders without formal authority. Negotiation and conflict-resolution skills, with the ability to reconcile competing stakeholder priorities and hold a defensible "no." Executive storytelling - distilling complex technical and operational detail into concise written and verbal narratives for senior leadership. Customer and technician empathy, intellectual curiosity, and the adaptability to stay effective as priorities and requirements shift. Ability to manage multiple concurrent product initiatives without losing momentum on any of them. Benefits: Brinks Home recognizes the value of benefits for you and your family, so we offer a comprehensive and competitive benefits program: Medical, Dental, Vision, 401(k) with Employer Match, Paid Time Off & Paid Holidays, HSA/FSA, Life & AD&D Insurance, Disability Coverage, Maternity/Parental Leave, Mental & Physical Health Benefits, Employee Resource Groups, Volunteer Hours, Discounted Equipment & Monitoring, and Employee Referral Program To learn more about our company culture and career opportunities, please visit our LinkedIn and Career Page . Brinks Home provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local laws.
09/19/2026
Full time
Description Brinks Home is a leader in the smart security industry, protecting over one million people across the U.S., Canada, and Puerto Rico. Our platinum-grade protection is backed by award-winning customer service and expertly trained professionals. We strive for the highest standards for our customers while fostering a positive work environment for our employees. We create a culture that fosters innovation, celebrates creativity, and encourages authenticity. Join us and be part of a collaborative team that is relentless in our pursuit of security for life. Position Overview: We are currently seeking a seasoned, results-driven Technical Product Owner who embodies our core values: Service, Accountability, Customer Focus, Growth, and Integrity. This role will lead the strategy, development, and optimization of our sales and field service experiences across mobile and web - the tools our field technicians rely on to install, activate, and service equipment in customers' homes every day. You will ensure our products deliver exceptional value to users, with a focus on usability, performance, and innovation. You will operate as the Product Owner on a Scrum team within a SAFe environment, working shoulder-to-shoulder with Field Operations, Engineering, and Sales, and using AI to work faster and smarter across the entire product lifecycle. As a senior individual contributor, you will set direction in ambiguous problem spaces with minimal oversight, influence without authority, and raise the bar for product practice across the team. Key Responsibilities: Own and manage the mobile and web product backlog, equipment inventory capabilities, and scheduling and dispatch capabilities. Prioritize features that enhance the user experience across both platforms. Define clear user stories, acceptance criteria, and success metrics with instrumentation defined up front, then use experimentation - A/B tests, feature flags, and phased rollouts - and post-release data to measure adoption and iterate. Collaborate with UX/UI designers to ensure intuitive, responsive, and accessible mobile and web interfaces. Advocate for mobile-first, responsive design principles across the organization and stay current on mobile and web trends, technologies, and best practices. Partner with Field Operations leadership and field technicians to translate equipment job installation, activation, and service workflows into clear product requirements. Own and communicate the product roadmap for your area, sequencing releases against business outcomes, keeping stakeholders current on status and expectations, and making trade-offs visible to leadership. Identify, pilot, and champion AI tools that accelerate discovery, backlog refinement, documentation, and day-to-day team productivity. Define requirements for AI-enabled product features such as guided troubleshooting, intelligent scheduling, and automated job verification - including data quality, human-in-the-loop review, and responsible-AI guardrails. Serve as Product Owner on a Scrum team operating within a SAFe cadence - PI planning, backlog refinement, iteration planning, reviews, and system demos. Manage cross-team dependencies and integration requirements across CRM, dispatch and scheduling, billing, e-signature, order management and inventory, and device telemetry systems. Partner with developers across the full software development lifecycle, from discovery through release - technical feasibility, timely delivery, QA, UAT, and deployment readiness. Define non-functional requirements with engineering - performance, offline and low-connectivity resilience, security, accessibility, and data-privacy compliance. Drive alignment across Field Ops, Sales, Engineering, QA, Training, and Customer Care in a fast-paced, cross-functional environment. Own release readiness and adoption: release notes, field enablement and training content, pilot and rollout plans, and post-launch issue triage with engineering. Required Qualifications: 7+ years of experience as a Product Owner, with a strong focus on mobile and web applications, including at least 3 years owning a technical product alongside an engineering team. 5+ years of experience with SAFe Agile methodology and Jira. AI-savvy - strong understanding of AI/ML for identifying opportunities, requirements, and design aligned with business goals, plus practical use of AI in day-to-day work to maximize individual and team productivity. Hands-on experience as a Product Owner on a Scrum team, including representing a product area across multiple PI planning cycles; SAFe (Scaled Agile Framework) experience strongly preferred. Command of the software development lifecycle and its supporting functions, including QA, release management, DevOps, documentation, and training. Strong requirements craft - user story mapping, process and workflow mapping, clear written specifications, and testable acceptance criteria. Proven track record of delivering successful, high-usage mobile (iOS and/or Android) and web applications from concept through multiple release cycles. Data fluency - comfortable querying data or working in BI dashboards, and using funnel, cohort, and adoption analysis to size opportunities and defend decisions. Ability to operate autonomously - setting direction in ambiguity, making defensible trade-off decisions, and driving outcomes end to end. Excellent communication, stakeholder management, and problem-solving skills, with the ability to communicate effectively across all levels of operations and users. Proven ability to thrive on fast-paced, cross-functional teams with competing priorities and shifting requirements. Preferred Qualifications: Experience within the home security industry and/or a field operations dispatch. Working knowledge of scheduling capabilities and inventory management. Experience with A/B testing and product analytics across mobile and web. Proficiency with backlog management, documentation, and design collaboration tooling. Experience mentoring or coaching less-experienced product owners and analysts, and influencing senior stakeholders without formal authority. Negotiation and conflict-resolution skills, with the ability to reconcile competing stakeholder priorities and hold a defensible "no." Executive storytelling - distilling complex technical and operational detail into concise written and verbal narratives for senior leadership. Customer and technician empathy, intellectual curiosity, and the adaptability to stay effective as priorities and requirements shift. Ability to manage multiple concurrent product initiatives without losing momentum on any of them. Benefits: Brinks Home recognizes the value of benefits for you and your family, so we offer a comprehensive and competitive benefits program: Medical, Dental, Vision, 401(k) with Employer Match, Paid Time Off & Paid Holidays, HSA/FSA, Life & AD&D Insurance, Disability Coverage, Maternity/Parental Leave, Mental & Physical Health Benefits, Employee Resource Groups, Volunteer Hours, Discounted Equipment & Monitoring, and Employee Referral Program To learn more about our company culture and career opportunities, please visit our LinkedIn and Career Page . Brinks Home provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local laws.
Manager of Software Engineering - Secret Clearance
Amarx Search, Inc. Pennsauken, New Jersey
Direct Hire - Full Time position in Camden, NJ Position ID: 2762 An excellent position with a large defense technology company delivering innovative mission solutions Manager of Software Engineering - Secret Clearance required Please apply ONLY if you have a DoD Secret clearance and 9 years relevant experience United States Citizenship is required due to government contract requirements We can ONLY consider your application if you have: 1: Bachelor's Degree 2: 9+ years relevant experience (7 with Master's, 13 with no degree) 3: Experience in a senior engineering, technical lead, or informal leadership role 4: Demonstrated experience designing, developing, or integrating RESTful services in a production environment. 5: DoD Secret security clearance 6: Strong communication skills and the ability to collaborate across disciplines and programs. 7: Exposure to proposal development, cost estimation, or EVMS-driven program environments. 8: Experience working in an Agile or hybrid Agile development environment. The company's Space Mission Systems is seeking an experienced Software Engineering Manager to lead development of mission-critical software, supporting communications and service-based systems. The Software Engineering Manager will collaborate across engineering disciplines, mentor engineers, support recruiting and performance management activities, and contribute technical input to proposal efforts and engineering planning. The selected candidate will guide software development across the full lifecycle including requirements analysis, architecture, design, implementation, integration, and test. This role provides day-to-day technical leadership for a team of software engineers while ensuring delivery of high-quality software aligned with cost, schedule, and technical performance objectives. The successful candidate will also support development and sustainment of VoIP and real-time communications applications utilizing SIP-based call control and frameworks such as PJSIP or PJMEDIA. DESIRED (not required) SKILLS: Hands-on experience deploying or sustaining applications in OpenShift, MicroShift, or Kubernetes-based environments. Experience with DevOps practices and tools, including CI/CD pipelines, automated testing, and configuration management. Experience with Linux-based development, specifically Red Hat Enterprise Linux (RHEL). Experience building, deploying, or sustaining software packaged as RPMs. Experience with Java enterprise applications, including WAR-based deployments. Familiarity with JBoss or similar Java application servers. Experience or exposure to VoIP systems, SIP-based call control, or real-time media applications. Familiarity with PJSIP / PJMEDIA or similar VoIP/media frameworks. Duties and Responsibilities Provide day-to-day technical and task leadership for a team of software engineers, ensuring delivery of high-quality software on cost, schedule, and technical performance. Will lead development of RESTful APIs and service-based architectures and supporting Java enterprise applications deployed as WAR files on JBoss or similar application servers. Will guide software deployment and sustainment on Red Hat Enterprise Linux systems and supporting RPM-based software packaging and distribution. Participate in and lead code reviews, design reviews, and architecture discussions for service-based, VoIP, and application-level software systems. Collaborate closely with systems engineering, electrical engineering, cybersecurity, and program leadership to ensure successful end-to-end system integration. Mentor and coach engineers on technical growth, career development, and performance expectations. Support performance management activities, including providing input to performance reviews and goal setting. Assist with recruiting, interviewing, onboarding, and tasking of software engineers. Maintain hands-on technical involvement as required to support development, integration, troubleshooting, and sustainment activities. Provide technical input to proposal efforts, including rough-order-of-magnitude estimates, task definition, and engineering assumptions. Please send resume to - Amarx Search, Inc. - (link removed)
09/16/2026
Direct Hire - Full Time position in Camden, NJ Position ID: 2762 An excellent position with a large defense technology company delivering innovative mission solutions Manager of Software Engineering - Secret Clearance required Please apply ONLY if you have a DoD Secret clearance and 9 years relevant experience United States Citizenship is required due to government contract requirements We can ONLY consider your application if you have: 1: Bachelor's Degree 2: 9+ years relevant experience (7 with Master's, 13 with no degree) 3: Experience in a senior engineering, technical lead, or informal leadership role 4: Demonstrated experience designing, developing, or integrating RESTful services in a production environment. 5: DoD Secret security clearance 6: Strong communication skills and the ability to collaborate across disciplines and programs. 7: Exposure to proposal development, cost estimation, or EVMS-driven program environments. 8: Experience working in an Agile or hybrid Agile development environment. The company's Space Mission Systems is seeking an experienced Software Engineering Manager to lead development of mission-critical software, supporting communications and service-based systems. The Software Engineering Manager will collaborate across engineering disciplines, mentor engineers, support recruiting and performance management activities, and contribute technical input to proposal efforts and engineering planning. The selected candidate will guide software development across the full lifecycle including requirements analysis, architecture, design, implementation, integration, and test. This role provides day-to-day technical leadership for a team of software engineers while ensuring delivery of high-quality software aligned with cost, schedule, and technical performance objectives. The successful candidate will also support development and sustainment of VoIP and real-time communications applications utilizing SIP-based call control and frameworks such as PJSIP or PJMEDIA. DESIRED (not required) SKILLS: Hands-on experience deploying or sustaining applications in OpenShift, MicroShift, or Kubernetes-based environments. Experience with DevOps practices and tools, including CI/CD pipelines, automated testing, and configuration management. Experience with Linux-based development, specifically Red Hat Enterprise Linux (RHEL). Experience building, deploying, or sustaining software packaged as RPMs. Experience with Java enterprise applications, including WAR-based deployments. Familiarity with JBoss or similar Java application servers. Experience or exposure to VoIP systems, SIP-based call control, or real-time media applications. Familiarity with PJSIP / PJMEDIA or similar VoIP/media frameworks. Duties and Responsibilities Provide day-to-day technical and task leadership for a team of software engineers, ensuring delivery of high-quality software on cost, schedule, and technical performance. Will lead development of RESTful APIs and service-based architectures and supporting Java enterprise applications deployed as WAR files on JBoss or similar application servers. Will guide software deployment and sustainment on Red Hat Enterprise Linux systems and supporting RPM-based software packaging and distribution. Participate in and lead code reviews, design reviews, and architecture discussions for service-based, VoIP, and application-level software systems. Collaborate closely with systems engineering, electrical engineering, cybersecurity, and program leadership to ensure successful end-to-end system integration. Mentor and coach engineers on technical growth, career development, and performance expectations. Support performance management activities, including providing input to performance reviews and goal setting. Assist with recruiting, interviewing, onboarding, and tasking of software engineers. Maintain hands-on technical involvement as required to support development, integration, troubleshooting, and sustainment activities. Provide technical input to proposal efforts, including rough-order-of-magnitude estimates, task definition, and engineering assumptions. Please send resume to - Amarx Search, Inc. - (link removed)
Principal Staff Engineer - Web Platform
PlanetArt Agoura Hills, California
Job Description Job Description Company and Vision PlanetArt's vision is to be the leading seller of personalized and make-on-demand products worldwide. We provide consumers with unmatched tools and content and an unparalleled end-to-end customer experience that result in high-quality, meaningful finished products and memorable celebrations of live events. The company's brands include the popular FreePrints and FreePrints Photobooks apps and the industry leading SimplytoImpress card and stationery site, as well as Personal Creations, CafePress and ISeeMe! Visit to learn more about our brands. We have more than 500 team members across multiple offices, primarily in Calabasas CA, San Diego CA, Woodridge IL, Minneapolis, MN and Pleasanton, CA. We also have team members in two company-owned offices in China, as well as in Europe. Job Overview PlanetArt is seeking a Principal Staff Engineer-Web Platform to serve as a senior technical leader within our engineering organization and a key partner to the VP of Engineering. This is a highly hands-on role focused on building, operating, and scaling high-traffic ecommerce web platforms in a fast-moving production environment. Reporting directly to the VP of Engineering, this engineer will play a critical role in architecting, developing, troubleshooting, and maintaining our LAMP-based web applications and AWS infrastructure. The ideal candidate is equally comfortable writing production code, diagnosing complex site reliability issues, managing cloud infrastructure, and leading technical problem-solving during high-severity incidents. This role requires strong operational judgment and the ability to independently own production challenges across application, database, infrastructure, and deployment layers. The engineer will collaborate closely with our China-based development organization, serving as a senior US-based technical lead responsible for cross-team coordination, code quality, architectural guidance, and production stability. This is an ideal opportunity for an experienced engineer who thrives in high-scale ecommerce environments and enjoys combining deep application engineering with modern cloud operations and production ownership. PLEASE NOTE: Candidates much be local to or willing to relocate to the Calabasas area as we operate on a hybrid work model (3 days onsite, 2 remote) What You'll Do Key Responsibilities Full-Stack Platform Engineering: Design, develop, and maintain scalable features and services across our LAMP-based ecommerce platform, with a strong focus on reliability, performance, maintainability, and operational excellence. Production Operations & Incident Response : Act as a senior technical escalation point for complex production incidents, troubleshooting issues across application, infrastructure, networking, database, CDN, and deployment layers. Lead root cause analysis and drive long-term stability improvements. AWS Infrastructure Ownership : Manage and optimize AWS infrastructure, including deployment architecture, scaling strategies, observability, security, disaster recovery, and cost efficiency. Partner closely with DevOps and engineering leadership on operational best practices. High-Scale Performance Optimization : Monitor and improve application, database, and infrastructure performance for high-traffic consumer web applications. Identify bottlenecks and implement scalable solutions to improve uptime, latency, and system resilience. Cross-Functional Technical Leadership: Partner with Product, Design, Operations, and Customer Experience teams to translate business requirements into scalable technical solutions and ensure successful project execution. Global Engineering Collaboration : Work closely with the China-based engineering team to coordinate development efforts, conduct code reviews, align on architectural direction, manage releases, and maintain strong engineering communication across time zones. Code Quality & Engineering Standards : Champion high engineering standards through code reviews, testing strategies, documentation, observability, and operational best practices. Drive continuous improvement in system reliability and development processes. Technical Mentorship & Leadership : Provide technical mentorship and architectural guidance across the engineering organization. Influence technical direction through hands-on leadership, strong execution, and collaborative problem-solving. Requirements What You Should Have Skills, Qualifications, and Requirements Senior-Level Full-Stack Engineering Experience: 5+ years of professional experience building and operating large-scale web applications, including substantial hands-on experience with the LAMP stack (Linux, Apache, MySQL, PHP). Strong AWS & Cloud Operations Expertise : Deep hands-on experience with AWS services and production cloud environments, including EC2, RDS, S3, Lambda, CloudWatch, networking, scaling, monitoring, and infrastructure troubleshooting. Ecommerce & High-Traffic Website Experience : Experience supporting high-volume consumer-facing websites or ecommerce platforms, with a strong understanding of scalability, uptime, performance optimization, and operational reliability. Production Troubleshooting Expertise : Demonstrated ability to diagnose and resolve complex production issues under pressure, including database replication issues, performance degradation, infrastructure failures, deployment issues, and site outages. Distributed Systems & Database Knowledge : Strong understanding of distributed web architectures, database performance tuning, replication strategies, caching, queuing systems, and fault-tolerant system design. Global Team Collaboration: Experience working effectively with offshore or globally distributed engineering teams, with strong communication, coordination, and cross-cultural collaboration skills. Chinese Language Skills: Ability to communicate in Mandarin (spoken or written) is highly desirable to facilitate collaboration with our China-based engineering team. Engineering Best Practices: Strong understanding of software engineering fundamentals including Git workflows, CI/CD pipelines, automated testing, observability, code review practices, and secure development standards. Ownership Mentality: Self-directed engineer with strong operational instincts, excellent judgment, and the ability to independently own critical technical initiatives from design through production support. Technical Leadership: Demonstrated ability to influence engineering direction, mentor developers, and drive technical excellence through hands-on leadership rather than direct people management. What You Can Expect Working Conditions Work is performed in an office environment with low to moderate noise levels. Position requires regular, continuous use of computer. Position requires regular sitting and standing. Position requires regular interaction with team members through the following methods: in-person, phone, Zoom, Slack, or email. May require occasional travel. This is a hybrid position; employees are expected to be in the office three days per week (Monday, Tuesday, and Thursday) with the option of working remotely two days (Wednesday and Friday). Benefits The compensation range for this position is $130,000-$220,000 annual salary. PlanetArt offers a comprehensive benefits package, including: Health, Dental, and Vision Insurance Life Insurance Pet Insurance Mental Health Insurance 401(k) with matching Comprehensive Time Off Program including Paid Time Off, Sick Days, Paid Holidays, and Floating Holidays Employee Product Discounts
09/15/2026
Full time
Job Description Job Description Company and Vision PlanetArt's vision is to be the leading seller of personalized and make-on-demand products worldwide. We provide consumers with unmatched tools and content and an unparalleled end-to-end customer experience that result in high-quality, meaningful finished products and memorable celebrations of live events. The company's brands include the popular FreePrints and FreePrints Photobooks apps and the industry leading SimplytoImpress card and stationery site, as well as Personal Creations, CafePress and ISeeMe! Visit to learn more about our brands. We have more than 500 team members across multiple offices, primarily in Calabasas CA, San Diego CA, Woodridge IL, Minneapolis, MN and Pleasanton, CA. We also have team members in two company-owned offices in China, as well as in Europe. Job Overview PlanetArt is seeking a Principal Staff Engineer-Web Platform to serve as a senior technical leader within our engineering organization and a key partner to the VP of Engineering. This is a highly hands-on role focused on building, operating, and scaling high-traffic ecommerce web platforms in a fast-moving production environment. Reporting directly to the VP of Engineering, this engineer will play a critical role in architecting, developing, troubleshooting, and maintaining our LAMP-based web applications and AWS infrastructure. The ideal candidate is equally comfortable writing production code, diagnosing complex site reliability issues, managing cloud infrastructure, and leading technical problem-solving during high-severity incidents. This role requires strong operational judgment and the ability to independently own production challenges across application, database, infrastructure, and deployment layers. The engineer will collaborate closely with our China-based development organization, serving as a senior US-based technical lead responsible for cross-team coordination, code quality, architectural guidance, and production stability. This is an ideal opportunity for an experienced engineer who thrives in high-scale ecommerce environments and enjoys combining deep application engineering with modern cloud operations and production ownership. PLEASE NOTE: Candidates much be local to or willing to relocate to the Calabasas area as we operate on a hybrid work model (3 days onsite, 2 remote) What You'll Do Key Responsibilities Full-Stack Platform Engineering: Design, develop, and maintain scalable features and services across our LAMP-based ecommerce platform, with a strong focus on reliability, performance, maintainability, and operational excellence. Production Operations & Incident Response : Act as a senior technical escalation point for complex production incidents, troubleshooting issues across application, infrastructure, networking, database, CDN, and deployment layers. Lead root cause analysis and drive long-term stability improvements. AWS Infrastructure Ownership : Manage and optimize AWS infrastructure, including deployment architecture, scaling strategies, observability, security, disaster recovery, and cost efficiency. Partner closely with DevOps and engineering leadership on operational best practices. High-Scale Performance Optimization : Monitor and improve application, database, and infrastructure performance for high-traffic consumer web applications. Identify bottlenecks and implement scalable solutions to improve uptime, latency, and system resilience. Cross-Functional Technical Leadership: Partner with Product, Design, Operations, and Customer Experience teams to translate business requirements into scalable technical solutions and ensure successful project execution. Global Engineering Collaboration : Work closely with the China-based engineering team to coordinate development efforts, conduct code reviews, align on architectural direction, manage releases, and maintain strong engineering communication across time zones. Code Quality & Engineering Standards : Champion high engineering standards through code reviews, testing strategies, documentation, observability, and operational best practices. Drive continuous improvement in system reliability and development processes. Technical Mentorship & Leadership : Provide technical mentorship and architectural guidance across the engineering organization. Influence technical direction through hands-on leadership, strong execution, and collaborative problem-solving. Requirements What You Should Have Skills, Qualifications, and Requirements Senior-Level Full-Stack Engineering Experience: 5+ years of professional experience building and operating large-scale web applications, including substantial hands-on experience with the LAMP stack (Linux, Apache, MySQL, PHP). Strong AWS & Cloud Operations Expertise : Deep hands-on experience with AWS services and production cloud environments, including EC2, RDS, S3, Lambda, CloudWatch, networking, scaling, monitoring, and infrastructure troubleshooting. Ecommerce & High-Traffic Website Experience : Experience supporting high-volume consumer-facing websites or ecommerce platforms, with a strong understanding of scalability, uptime, performance optimization, and operational reliability. Production Troubleshooting Expertise : Demonstrated ability to diagnose and resolve complex production issues under pressure, including database replication issues, performance degradation, infrastructure failures, deployment issues, and site outages. Distributed Systems & Database Knowledge : Strong understanding of distributed web architectures, database performance tuning, replication strategies, caching, queuing systems, and fault-tolerant system design. Global Team Collaboration: Experience working effectively with offshore or globally distributed engineering teams, with strong communication, coordination, and cross-cultural collaboration skills. Chinese Language Skills: Ability to communicate in Mandarin (spoken or written) is highly desirable to facilitate collaboration with our China-based engineering team. Engineering Best Practices: Strong understanding of software engineering fundamentals including Git workflows, CI/CD pipelines, automated testing, observability, code review practices, and secure development standards. Ownership Mentality: Self-directed engineer with strong operational instincts, excellent judgment, and the ability to independently own critical technical initiatives from design through production support. Technical Leadership: Demonstrated ability to influence engineering direction, mentor developers, and drive technical excellence through hands-on leadership rather than direct people management. What You Can Expect Working Conditions Work is performed in an office environment with low to moderate noise levels. Position requires regular, continuous use of computer. Position requires regular sitting and standing. Position requires regular interaction with team members through the following methods: in-person, phone, Zoom, Slack, or email. May require occasional travel. This is a hybrid position; employees are expected to be in the office three days per week (Monday, Tuesday, and Thursday) with the option of working remotely two days (Wednesday and Friday). Benefits The compensation range for this position is $130,000-$220,000 annual salary. PlanetArt offers a comprehensive benefits package, including: Health, Dental, and Vision Insurance Life Insurance Pet Insurance Mental Health Insurance 401(k) with matching Comprehensive Time Off Program including Paid Time Off, Sick Days, Paid Holidays, and Floating Holidays Employee Product Discounts
Senior MLOps Engineer - Snowflake
KAPI LLC Addison, Texas
Job Description Job Description Work Arrangement: Dallas-based / Hybrid Visa Sponsorship: Not available. Candidates must already be authorized to work in the United States without current or future employer sponsorship. Job Summary We are seeking a highly experienced Senior MLOps Engineer with strong hands-on Snowflake experience to support enterprise machine learning platforms and production ML workloads. The ideal candidate has hands-on experience taking machine learning models from experimentation through production and building the deployment pipelines, monitoring, automation, infrastructure, and governance capabilities required to operate ML solutions reliably at enterprise scale. This role will work closely with Data Scientists, ML Engineers, Data Engineers, Cloud Engineers, and enterprise platform teams. Key Responsibilities Design, build, and maintain enterprise-grade MLOps platforms and pipelines. Operationalize machine learning models developed by Data Science teams. Build automated ML workflows covering training, validation, deployment, monitoring, retraining, and retirement. Implement CI/CD pipelines specifically for machine learning workloads. Establish model registry, versioning, lineage, artifact management, and reproducibility. Implement model monitoring, data drift detection, model drift detection, prediction-quality monitoring, and alerting. Integrate ML workloads with Snowflake-based enterprise data environments . Build and optimize Python- and SQL-based data and ML pipelines. Support Snowflake data ingestion, transformation, compute, security, and ML integrations. Containerize ML workloads using Docker and deploy workloads through Kubernetes or comparable orchestration platforms. Implement logging, observability, alerting, and production support processes. Automate deployment and infrastructure provisioning using modern DevOps and Infrastructure-as-Code practices. Support model governance, approval workflows, lineage, auditability, and access controls. Troubleshoot production ML pipelines, model-serving infrastructure, Snowflake integrations, and performance issues. Develop reusable MLOps frameworks, standards, templates, and best practices. Mandatory Qualifications Candidates must have hands-on production experience in both MLOps and Snowflake . MLOps - Required Strong production experience with: ML model deployment and operationalization Model lifecycle management ML CI/CD Experiment tracking Model registry and versioning Automated model validation Model monitoring Data and model drift detection Retraining pipelines Pipeline orchestration Production troubleshooting Experience with one or more of the following: ML flow Kubeflow AWS SageMaker Azure Machine Learning Airflow Argo Workflows Prefect Dagster Equivalent enterprise MLOps platforms Snowflake - Required Strong hands-on Snowflake experience including: Snowflake architecture Databases, schemas, tables, and views Virtual warehouses Compute management Snowflake security and RBAC Data ingestion and transformation Performance optimization Python integration Snowflake integration with ML pipelines Experience with the following is strongly preferred: Snowpark Snowpark Python Snowflake ML Snowflake Model Registry Snowflake Feature Store Snowflake Tasks and Streams Dynamic Tables Snowpipe Cortex / Snowflake AI capabilities Additional Required Technical Skills Strong Python Strong SQL Git REST APIs Linux Shell scripting Docker CI/CD Cloud platforms such as AWS, Azure, or GCP Preferred Skills Experience with: Kubernetes Terraform GitHub Actions Jenkins GitLab CI/CD Azure DevOps dbt Spark Kafka Grafana CloudWatch Evidently Education and Experience Bachelor's or Master's degree in Computer Science, Engineering, Data Science, Machine Learning, or related field. 6+ years of software, cloud, data, or ML engineering experience. 3+ years of hands-on production MLOps experience. Strong hands-on Snowflake experience. Experience deploying ML models into production. Experience implementing ML CI/CD pipelines. Strong Python and SQL skills. Experience with Docker and cloud infrastructure. Work Authorization This position does not provide visa sponsorship. Candidates must be currently authorized to work in the United States without employer sponsorship and must not require sponsorship now or in the future. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it.
09/15/2026
Full time
Job Description Job Description Work Arrangement: Dallas-based / Hybrid Visa Sponsorship: Not available. Candidates must already be authorized to work in the United States without current or future employer sponsorship. Job Summary We are seeking a highly experienced Senior MLOps Engineer with strong hands-on Snowflake experience to support enterprise machine learning platforms and production ML workloads. The ideal candidate has hands-on experience taking machine learning models from experimentation through production and building the deployment pipelines, monitoring, automation, infrastructure, and governance capabilities required to operate ML solutions reliably at enterprise scale. This role will work closely with Data Scientists, ML Engineers, Data Engineers, Cloud Engineers, and enterprise platform teams. Key Responsibilities Design, build, and maintain enterprise-grade MLOps platforms and pipelines. Operationalize machine learning models developed by Data Science teams. Build automated ML workflows covering training, validation, deployment, monitoring, retraining, and retirement. Implement CI/CD pipelines specifically for machine learning workloads. Establish model registry, versioning, lineage, artifact management, and reproducibility. Implement model monitoring, data drift detection, model drift detection, prediction-quality monitoring, and alerting. Integrate ML workloads with Snowflake-based enterprise data environments . Build and optimize Python- and SQL-based data and ML pipelines. Support Snowflake data ingestion, transformation, compute, security, and ML integrations. Containerize ML workloads using Docker and deploy workloads through Kubernetes or comparable orchestration platforms. Implement logging, observability, alerting, and production support processes. Automate deployment and infrastructure provisioning using modern DevOps and Infrastructure-as-Code practices. Support model governance, approval workflows, lineage, auditability, and access controls. Troubleshoot production ML pipelines, model-serving infrastructure, Snowflake integrations, and performance issues. Develop reusable MLOps frameworks, standards, templates, and best practices. Mandatory Qualifications Candidates must have hands-on production experience in both MLOps and Snowflake . MLOps - Required Strong production experience with: ML model deployment and operationalization Model lifecycle management ML CI/CD Experiment tracking Model registry and versioning Automated model validation Model monitoring Data and model drift detection Retraining pipelines Pipeline orchestration Production troubleshooting Experience with one or more of the following: ML flow Kubeflow AWS SageMaker Azure Machine Learning Airflow Argo Workflows Prefect Dagster Equivalent enterprise MLOps platforms Snowflake - Required Strong hands-on Snowflake experience including: Snowflake architecture Databases, schemas, tables, and views Virtual warehouses Compute management Snowflake security and RBAC Data ingestion and transformation Performance optimization Python integration Snowflake integration with ML pipelines Experience with the following is strongly preferred: Snowpark Snowpark Python Snowflake ML Snowflake Model Registry Snowflake Feature Store Snowflake Tasks and Streams Dynamic Tables Snowpipe Cortex / Snowflake AI capabilities Additional Required Technical Skills Strong Python Strong SQL Git REST APIs Linux Shell scripting Docker CI/CD Cloud platforms such as AWS, Azure, or GCP Preferred Skills Experience with: Kubernetes Terraform GitHub Actions Jenkins GitLab CI/CD Azure DevOps dbt Spark Kafka Grafana CloudWatch Evidently Education and Experience Bachelor's or Master's degree in Computer Science, Engineering, Data Science, Machine Learning, or related field. 6+ years of software, cloud, data, or ML engineering experience. 3+ years of hands-on production MLOps experience. Strong hands-on Snowflake experience. Experience deploying ML models into production. Experience implementing ML CI/CD pipelines. Strong Python and SQL skills. Experience with Docker and cloud infrastructure. Work Authorization This position does not provide visa sponsorship. Candidates must be currently authorized to work in the United States without employer sponsorship and must not require sponsorship now or in the future. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it.
Senior MLOps Engineer - Snowflake
KAPI LLC Addison, Texas
Job Description Job Description Work Arrangement: Dallas-based / Hybrid Visa Sponsorship: Not available. Candidates must already be authorized to work in the United States without current or future employer sponsorship. Job Summary We are seeking a highly experienced Senior MLOps Engineer with strong hands-on Snowflake experience to support enterprise machine learning platforms and production ML workloads. The ideal candidate has hands-on experience taking machine learning models from experimentation through production and building the deployment pipelines, monitoring, automation, infrastructure, and governance capabilities required to operate ML solutions reliably at enterprise scale. This role will work closely with Data Scientists, ML Engineers, Data Engineers, Cloud Engineers, and enterprise platform teams. Key Responsibilities Design, build, and maintain enterprise-grade MLOps platforms and pipelines. Operationalize machine learning models developed by Data Science teams. Build automated ML workflows covering training, validation, deployment, monitoring, retraining, and retirement. Implement CI/CD pipelines specifically for machine learning workloads. Establish model registry, versioning, lineage, artifact management, and reproducibility. Implement model monitoring, data drift detection, model drift detection, prediction-quality monitoring, and alerting. Integrate ML workloads with Snowflake-based enterprise data environments . Build and optimize Python- and SQL-based data and ML pipelines. Support Snowflake data ingestion, transformation, compute, security, and ML integrations. Containerize ML workloads using Docker and deploy workloads through Kubernetes or comparable orchestration platforms. Implement logging, observability, alerting, and production support processes. Automate deployment and infrastructure provisioning using modern DevOps and Infrastructure-as-Code practices. Support model governance, approval workflows, lineage, auditability, and access controls. Troubleshoot production ML pipelines, model-serving infrastructure, Snowflake integrations, and performance issues. Develop reusable MLOps frameworks, standards, templates, and best practices. Mandatory Qualifications Candidates must have hands-on production experience in both MLOps and Snowflake . MLOps - Required Strong production experience with: ML model deployment and operationalization Model lifecycle management ML CI/CD Experiment tracking Model registry and versioning Automated model validation Model monitoring Data and model drift detection Retraining pipelines Pipeline orchestration Production troubleshooting Experience with one or more of the following: ML flow Kubeflow AWS SageMaker Azure Machine Learning Airflow Argo Workflows Prefect Dagster Equivalent enterprise MLOps platforms Snowflake - Required Strong hands-on Snowflake experience including: Snowflake architecture Databases, schemas, tables, and views Virtual warehouses Compute management Snowflake security and RBAC Data ingestion and transformation Performance optimization Python integration Snowflake integration with ML pipelines Experience with the following is strongly preferred: Snowpark Snowpark Python Snowflake ML Snowflake Model Registry Snowflake Feature Store Snowflake Tasks and Streams Dynamic Tables Snowpipe Cortex / Snowflake AI capabilities Additional Required Technical Skills Strong Python Strong SQL Git REST APIs Linux Shell scripting Docker CI/CD Cloud platforms such as AWS, Azure, or GCP Preferred Skills Experience with: Kubernetes Terraform GitHub Actions Jenkins GitLab CI/CD Azure DevOps dbt Spark Kafka Grafana CloudWatch Evidently Education and Experience Bachelor's or Master's degree in Computer Science, Engineering, Data Science, Machine Learning, or related field. 6+ years of software, cloud, data, or ML engineering experience. 3+ years of hands-on production MLOps experience. Strong hands-on Snowflake experience. Experience deploying ML models into production. Experience implementing ML CI/CD pipelines. Strong Python and SQL skills. Experience with Docker and cloud infrastructure. Work Authorization This position does not provide visa sponsorship. Candidates must be currently authorized to work in the United States without employer sponsorship and must not require sponsorship now or in the future. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it.
09/13/2026
Full time
Job Description Job Description Work Arrangement: Dallas-based / Hybrid Visa Sponsorship: Not available. Candidates must already be authorized to work in the United States without current or future employer sponsorship. Job Summary We are seeking a highly experienced Senior MLOps Engineer with strong hands-on Snowflake experience to support enterprise machine learning platforms and production ML workloads. The ideal candidate has hands-on experience taking machine learning models from experimentation through production and building the deployment pipelines, monitoring, automation, infrastructure, and governance capabilities required to operate ML solutions reliably at enterprise scale. This role will work closely with Data Scientists, ML Engineers, Data Engineers, Cloud Engineers, and enterprise platform teams. Key Responsibilities Design, build, and maintain enterprise-grade MLOps platforms and pipelines. Operationalize machine learning models developed by Data Science teams. Build automated ML workflows covering training, validation, deployment, monitoring, retraining, and retirement. Implement CI/CD pipelines specifically for machine learning workloads. Establish model registry, versioning, lineage, artifact management, and reproducibility. Implement model monitoring, data drift detection, model drift detection, prediction-quality monitoring, and alerting. Integrate ML workloads with Snowflake-based enterprise data environments . Build and optimize Python- and SQL-based data and ML pipelines. Support Snowflake data ingestion, transformation, compute, security, and ML integrations. Containerize ML workloads using Docker and deploy workloads through Kubernetes or comparable orchestration platforms. Implement logging, observability, alerting, and production support processes. Automate deployment and infrastructure provisioning using modern DevOps and Infrastructure-as-Code practices. Support model governance, approval workflows, lineage, auditability, and access controls. Troubleshoot production ML pipelines, model-serving infrastructure, Snowflake integrations, and performance issues. Develop reusable MLOps frameworks, standards, templates, and best practices. Mandatory Qualifications Candidates must have hands-on production experience in both MLOps and Snowflake . MLOps - Required Strong production experience with: ML model deployment and operationalization Model lifecycle management ML CI/CD Experiment tracking Model registry and versioning Automated model validation Model monitoring Data and model drift detection Retraining pipelines Pipeline orchestration Production troubleshooting Experience with one or more of the following: ML flow Kubeflow AWS SageMaker Azure Machine Learning Airflow Argo Workflows Prefect Dagster Equivalent enterprise MLOps platforms Snowflake - Required Strong hands-on Snowflake experience including: Snowflake architecture Databases, schemas, tables, and views Virtual warehouses Compute management Snowflake security and RBAC Data ingestion and transformation Performance optimization Python integration Snowflake integration with ML pipelines Experience with the following is strongly preferred: Snowpark Snowpark Python Snowflake ML Snowflake Model Registry Snowflake Feature Store Snowflake Tasks and Streams Dynamic Tables Snowpipe Cortex / Snowflake AI capabilities Additional Required Technical Skills Strong Python Strong SQL Git REST APIs Linux Shell scripting Docker CI/CD Cloud platforms such as AWS, Azure, or GCP Preferred Skills Experience with: Kubernetes Terraform GitHub Actions Jenkins GitLab CI/CD Azure DevOps dbt Spark Kafka Grafana CloudWatch Evidently Education and Experience Bachelor's or Master's degree in Computer Science, Engineering, Data Science, Machine Learning, or related field. 6+ years of software, cloud, data, or ML engineering experience. 3+ years of hands-on production MLOps experience. Strong hands-on Snowflake experience. Experience deploying ML models into production. Experience implementing ML CI/CD pipelines. Strong Python and SQL skills. Experience with Docker and cloud infrastructure. Work Authorization This position does not provide visa sponsorship. Candidates must be currently authorized to work in the United States without employer sponsorship and must not require sponsorship now or in the future. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it. Company Description About KAPI Advisors LLC KAPI Advisors LLC is a forward-thinking boutique firm at the intersection of AI, Machine Learning, Data Analytics, Generative AI, Mathematical Optimization, and Cloud Platform Engineering. We build highly personalized, cutting-edge solutions tailored to the specific needs of each client - empowering organizations with powerful tools to unlock new business opportunities, operational efficiencies, and platform reliability at scale. At KAPI Advisors, we are passionate about solving complex, real-world challenges through technology. Our capabilities span intelligent agent-based systems, advanced data analytics, optimization algorithms, and robust cloud infrastructure - all designed to work together seamlessly. Our offerings include Generative AI systems that enable automation and intelligent decision-making, ML models that surface actionable insights, Mathematical Optimization techniques for supply chain management, resource allocation, and logistics, and Site Reliability Engineering practices that ensure the platforms powering these solutions remain performant, resilient, and production-ready. We believe that great AI and data products are only as strong as the infrastructure beneath them. That's why our SRE and DevOps practice is central to everything we build - from designing observability frameworks and incident response playbooks, to implementing chaos engineering, disaster recovery, and developer productivity tooling that reduces toil and accelerates delivery. As a boutique firm, we offer the agility and depth that larger organizations simply cannot match. Our team works directly with clients to understand their unique challenges and craft tailored strategies that deliver both immediate impact and long-term resilience. Whether you're looking to scale an AI platform, modernize your cloud infrastructure, or build the reliability engineering culture your engineering team deserves - KAPI Advisors is the partner built for it.

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board