it job board logo
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
  • Recruiting? Post a job
  • Sign in
  • Sign up
  • Home
  • Find IT Jobs
  • Register CV
  • Register as Employer
  • Contact us
  • Career Advice
Sorry, that job is no longer available. Here are some results that may be similar to the job you were looking for.

117 jobs found

Email me jobs like this
Refine Search
Current Search
devops engineer senior
Senior Cloud Engineer
Gridware San Francisco, California
Job Description Job Description About Gridware Gridware is a San Francisco-based technology company dedicated to protecting and enhancing the electrical grid. We pioneered a groundbreaking new class of grid management called active grid response (AGR), focused on monitoring the electrical, physical, and environmental aspects of the grid that affect reliability and safety. Gridware's advanced Active Grid Response platform uses high-precision sensors to detect potential issues early, enabling proactive maintenance and fault mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the grid operates efficiently. The company is backed by climate-tech and Silicon Valley investors. For more information, please visit . Role Description We're scaling the deployment of critical infrastructure monitoring devices to detect real-world fault events that lead to wildfires. The platform you'll build and operate ingests millions of events per day from devices in the field, powers customer-facing dashboards and alerting, and supports the data science work that turns raw signals into grid intelligence. You will own AWS infrastructure, Kubernetes (EKS), CI/CD, and observability end-to-end, partnering with our Cloud Security team to keep the platform safe and compliant, and with backend, firmware, and data teams to keep them shipping fast. As an early member of the DevOps team, you'll have a direct hand in shaping how Gridware builds, deploys, and runs production systems for years to come. Responsibilities Design, build, and operate scalable, secure, and highly available cloud infrastructure across AWS. Own and evolve our Kubernetes platform, enabling reliable application deployment and operations through GitOps best practices. Build and maintain CI/CD systems that improve developer velocity, release quality, and operational reliability. Manage and optimize event-driven infrastructure powering high-volume telemetry and device data pipelines. Define and maintain Infrastructure as Code standards, ensuring consistency, repeatability, and scalability across environments. Develop and enhance observability, monitoring, and incident response capabilities to support reliable production operations. Partner closely with Security and Engineering teams to strengthen platform security, access management, and operational resilience. Troubleshoot complex production issues, drive root cause analysis, and turn lessons learned into automation, tooling, and operational improvements. Required Skills 5+ years of experience in DevOps, SRE, or Platform Engineering operating production AWS environments Deep expertise with Kubernetes (EKS preferred), GitOps workflows (Argo CD/Flux), and Infrastructure as Code (Terraform) Strong experience building and maintaining CI/CD pipelines, ideally with GitHub Actions Hands-on experience operating distributed systems and cloud-native platforms (e.g., Kafka/MSK) Solid understanding of networking, DNS, TLS, identity/access management, and cloud security best practices Experience with observability, monitoring, and logging tools such as Grafana, Prometheus, Loki, or similar Strong Linux, scripting, and troubleshooting skills with the ability to debug complex production issues end-to-end Bonus Skills Experience operating Apollo Router / GraphQL federation gateways in production. Experience operating Argo Workflows or similar Kubernetes-native job / pipeline runners in production. Familiarity with Databricks or ML Ops pipelines for data and model deployment. Experience designing, operating, and exercising Disaster Recovery (DR) environments, including cross-region replication, backups, and tested failover runbooks. Experience with Tailscale or other zero-trust networking tools. Experience supporting IoT / embedded fleets at scale, including secure device-to-cloud connectivity. Experience in high-growth startup environments where you must wear many hats. This describes the ideal candidate; many of us have picked up this expertise along the way. Even if you meet only part of this list, we encourage you to apply! Gridware Technologies Inc. is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to any characteristic protected by applicable federal, state, or local law. Benefits Health, Dental & Vision (Gold and Platinum with some providers plans fully covered) Paid parental leave Alternating day off (every other Monday) "Off the Grid", a two week per year paid break for all employees. Commuter allowance Company-paid training
08/05/2026
Full time
Job Description Job Description About Gridware Gridware is a San Francisco-based technology company dedicated to protecting and enhancing the electrical grid. We pioneered a groundbreaking new class of grid management called active grid response (AGR), focused on monitoring the electrical, physical, and environmental aspects of the grid that affect reliability and safety. Gridware's advanced Active Grid Response platform uses high-precision sensors to detect potential issues early, enabling proactive maintenance and fault mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the grid operates efficiently. The company is backed by climate-tech and Silicon Valley investors. For more information, please visit . Role Description We're scaling the deployment of critical infrastructure monitoring devices to detect real-world fault events that lead to wildfires. The platform you'll build and operate ingests millions of events per day from devices in the field, powers customer-facing dashboards and alerting, and supports the data science work that turns raw signals into grid intelligence. You will own AWS infrastructure, Kubernetes (EKS), CI/CD, and observability end-to-end, partnering with our Cloud Security team to keep the platform safe and compliant, and with backend, firmware, and data teams to keep them shipping fast. As an early member of the DevOps team, you'll have a direct hand in shaping how Gridware builds, deploys, and runs production systems for years to come. Responsibilities Design, build, and operate scalable, secure, and highly available cloud infrastructure across AWS. Own and evolve our Kubernetes platform, enabling reliable application deployment and operations through GitOps best practices. Build and maintain CI/CD systems that improve developer velocity, release quality, and operational reliability. Manage and optimize event-driven infrastructure powering high-volume telemetry and device data pipelines. Define and maintain Infrastructure as Code standards, ensuring consistency, repeatability, and scalability across environments. Develop and enhance observability, monitoring, and incident response capabilities to support reliable production operations. Partner closely with Security and Engineering teams to strengthen platform security, access management, and operational resilience. Troubleshoot complex production issues, drive root cause analysis, and turn lessons learned into automation, tooling, and operational improvements. Required Skills 5+ years of experience in DevOps, SRE, or Platform Engineering operating production AWS environments Deep expertise with Kubernetes (EKS preferred), GitOps workflows (Argo CD/Flux), and Infrastructure as Code (Terraform) Strong experience building and maintaining CI/CD pipelines, ideally with GitHub Actions Hands-on experience operating distributed systems and cloud-native platforms (e.g., Kafka/MSK) Solid understanding of networking, DNS, TLS, identity/access management, and cloud security best practices Experience with observability, monitoring, and logging tools such as Grafana, Prometheus, Loki, or similar Strong Linux, scripting, and troubleshooting skills with the ability to debug complex production issues end-to-end Bonus Skills Experience operating Apollo Router / GraphQL federation gateways in production. Experience operating Argo Workflows or similar Kubernetes-native job / pipeline runners in production. Familiarity with Databricks or ML Ops pipelines for data and model deployment. Experience designing, operating, and exercising Disaster Recovery (DR) environments, including cross-region replication, backups, and tested failover runbooks. Experience with Tailscale or other zero-trust networking tools. Experience supporting IoT / embedded fleets at scale, including secure device-to-cloud connectivity. Experience in high-growth startup environments where you must wear many hats. This describes the ideal candidate; many of us have picked up this expertise along the way. Even if you meet only part of this list, we encourage you to apply! Gridware Technologies Inc. is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to any characteristic protected by applicable federal, state, or local law. Benefits Health, Dental & Vision (Gold and Platinum with some providers plans fully covered) Paid parental leave Alternating day off (every other Monday) "Off the Grid", a two week per year paid break for all employees. Commuter allowance Company-paid training
AI Engineer
Teserac, Inc. Sunnyvale, California
Job Description Job Description About the Role Teserac is building neuron , a unified AI-native platform for data center observability, intelligence, and workflow automation. neuron processes real-time telemetry from thousands of sensors, meters, and control systems across heterogeneous environments - giving infrastructure owners the visibility to monitor, analyze, automate, and proactively manage power operations with full situational awareness. An embedded AI teammate serves as every operator's always-on co-pilot: detecting anomalies, correlating events, and surfacing recommendations 24/7. We are seeking an AI/ML Engineer who is excited to build intelligent systems at the intersection of applied AI and critical infrastructure. You will work across the full AI development lifecycle - from data pipelines and model integration to agentic orchestration, evaluation, and production support - collaborating closely with a small, fast-moving engineering team. This is not a research-only role, but research thinking matters here. You will be expected to read papers, stay ahead of the field, and bring ideas to the table - then build them into production systems. Who We Are Looking For We care more about how you think than how many years are on your resume. This role is open to both junior and senior candidates. What matters is: You are genuinely excited about AI and infrastructure - not just one of them You learn fast, go deep, and can hold your own in a technical debate You have the engineering fundamentals to ship reliable systems You are proactive, curious, and comfortable with a steep learning curve You want to work on something technically hard that actually matters in the physical world If you are early in your career but have strong fundamentals, a track record of self-directed learning, and a portfolio that shows you build things - we want to hear from you. What You Will Work On Multi-agent orchestration and LLM-driven triage workflows Time-series modeling for anomaly detection, failure prediction, and health forecasting on multivariate telemetry Retrieval-augmented knowledge systems for operations teams Data and ML pipelines - ingestion, ETL, and dataset construction Fine-tuning and post-training of language models for operational use cases AI observability, evaluation frameworks, and production performance benchmarking Responsibilities Design, develop, and maintain AI-powered applications and automation workflows Integrate and optimize LLM APIs for production use cases Build and refine retrieval and knowledge-augmentation pipelines Develop evaluation frameworks to benchmark AI system performance Implement monitoring, tracing, and debugging capabilities for AI systems Read and synthesize relevant research; bring ideas forward and debate them with the team Contribute to AI architecture decisions and production hardening Stay current with the rapidly evolving AI/ML landscape Requirements Required Degree in Computer Science, Machine Learning, Mathematics, or a related field - or equivalent demonstrated experience Strong proficiency in Python Solid software engineering fundamentals: testing, version control, CI/CD Experience working with LLM APIs in applied contexts Familiarity with agentic system concepts - tool/function-calling, agent frameworks Daily use of AI-assisted coding tools (Cursor, Copilot, Claude Code, etc.) Ability to read ML research papers and translate ideas into practical experiments Strong analytical thinking and clear communication - you can argue a position and update it when wrong Preferred Professional AI/ML engineering experience (any level) Experience building agentic systems using frameworks such as LangGraph or LangChain; MCP a plus Time-series modeling - forecasting and anomaly/failure prediction on multivariate data Experience fine-tuning or post-training language models PyTorch and/or model serving frameworks (e.g., vLLM) Experience building data and ML pipelines - ingestion, ETL, dataset construction Familiarity with cloud ML platforms, particularly GCP (Vertex AI) LLM evaluation and benchmarking: harness design and eval loop development Domain experience with data center or industrial telemetry, BMS/OT protocols (Niagara, BACnet/Modbus) Background in DevOps, distributed systems, or observability tooling Benefits Health Care Plan (Medical, Dental & Vision) Paid Time Off (Vacation, Sick & Public Holidays) Free Food & Snacks Stock Option Plan 401(k)
08/05/2026
Full time
Job Description Job Description About the Role Teserac is building neuron , a unified AI-native platform for data center observability, intelligence, and workflow automation. neuron processes real-time telemetry from thousands of sensors, meters, and control systems across heterogeneous environments - giving infrastructure owners the visibility to monitor, analyze, automate, and proactively manage power operations with full situational awareness. An embedded AI teammate serves as every operator's always-on co-pilot: detecting anomalies, correlating events, and surfacing recommendations 24/7. We are seeking an AI/ML Engineer who is excited to build intelligent systems at the intersection of applied AI and critical infrastructure. You will work across the full AI development lifecycle - from data pipelines and model integration to agentic orchestration, evaluation, and production support - collaborating closely with a small, fast-moving engineering team. This is not a research-only role, but research thinking matters here. You will be expected to read papers, stay ahead of the field, and bring ideas to the table - then build them into production systems. Who We Are Looking For We care more about how you think than how many years are on your resume. This role is open to both junior and senior candidates. What matters is: You are genuinely excited about AI and infrastructure - not just one of them You learn fast, go deep, and can hold your own in a technical debate You have the engineering fundamentals to ship reliable systems You are proactive, curious, and comfortable with a steep learning curve You want to work on something technically hard that actually matters in the physical world If you are early in your career but have strong fundamentals, a track record of self-directed learning, and a portfolio that shows you build things - we want to hear from you. What You Will Work On Multi-agent orchestration and LLM-driven triage workflows Time-series modeling for anomaly detection, failure prediction, and health forecasting on multivariate telemetry Retrieval-augmented knowledge systems for operations teams Data and ML pipelines - ingestion, ETL, and dataset construction Fine-tuning and post-training of language models for operational use cases AI observability, evaluation frameworks, and production performance benchmarking Responsibilities Design, develop, and maintain AI-powered applications and automation workflows Integrate and optimize LLM APIs for production use cases Build and refine retrieval and knowledge-augmentation pipelines Develop evaluation frameworks to benchmark AI system performance Implement monitoring, tracing, and debugging capabilities for AI systems Read and synthesize relevant research; bring ideas forward and debate them with the team Contribute to AI architecture decisions and production hardening Stay current with the rapidly evolving AI/ML landscape Requirements Required Degree in Computer Science, Machine Learning, Mathematics, or a related field - or equivalent demonstrated experience Strong proficiency in Python Solid software engineering fundamentals: testing, version control, CI/CD Experience working with LLM APIs in applied contexts Familiarity with agentic system concepts - tool/function-calling, agent frameworks Daily use of AI-assisted coding tools (Cursor, Copilot, Claude Code, etc.) Ability to read ML research papers and translate ideas into practical experiments Strong analytical thinking and clear communication - you can argue a position and update it when wrong Preferred Professional AI/ML engineering experience (any level) Experience building agentic systems using frameworks such as LangGraph or LangChain; MCP a plus Time-series modeling - forecasting and anomaly/failure prediction on multivariate data Experience fine-tuning or post-training language models PyTorch and/or model serving frameworks (e.g., vLLM) Experience building data and ML pipelines - ingestion, ETL, dataset construction Familiarity with cloud ML platforms, particularly GCP (Vertex AI) LLM evaluation and benchmarking: harness design and eval loop development Domain experience with data center or industrial telemetry, BMS/OT protocols (Niagara, BACnet/Modbus) Background in DevOps, distributed systems, or observability tooling Benefits Health Care Plan (Medical, Dental & Vision) Paid Time Off (Vacation, Sick & Public Holidays) Free Food & Snacks Stock Option Plan 401(k)
Senior Site Reliability Engineer- Palo Alto, the US
Kody Palo Alto, California
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Cross-Border Collaboration: Act as a key technical bridge between our US operations and international engineering hubs, leveraging bilingual communication to streamline complex technical alignment. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
08/05/2026
Full time
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Cross-Border Collaboration: Act as a key technical bridge between our US operations and international engineering hubs, leveraging bilingual communication to streamline complex technical alignment. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
Senior Site Reliability Engineer- Sunnyvale, CA, the US
Kody Sunnyvale, California
Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Lead a dynamic and innovative team in a very rapidly growing company. Competitive package. Collaborative, inclusive environment where your contributions are recognized and valued.
08/05/2026
Full time
Job Description Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Lead a dynamic and innovative team in a very rapidly growing company. Competitive package. Collaborative, inclusive environment where your contributions are recognized and valued.
Senior Engineer, Internal Tools, Artificial Intelligence (AI) Required, Work From Home
Ginas Tech Jobs San Francisco, California
Job Description Job Description Job Description Senior Engineer, Internal Tools, Artificial Intelligence (AI) Required, Work From Home The Senior AI Engineer is the engineering backbone of the internal tools team. Building and maintaining the platforms that every team in the company relies on. Your work directly increases organizational efficiency and enables teams to move faster. The Senior AI Engineer will own systems end-to-end, from scoping and architecture through production deployment and iteration, connecting multiple business systems into a seamless, reliable internal ecosystem. This position is 100% Remote. Senior Engineer Responsibilities: Build & Ship: - Design, build, and maintain internal platforms and tools that serve People, Finance, Ops, Sales, and Engineering teams. - Own features, end-to-end requirements, architecture, implementation, testing, deployment, and monitoring. - Write clean, well-tested, production-grade code. You hold yourself to the same bar as customer-facing products. Architecture & Integration: - Build API-first integrations across the internal ecosystem connecting HRIS, CRM, finance platforms, knowledge management, and developer tools into a coherent stack. - Design for reliability, performance, and scale what you build today must hold as the company grows 5-10x. - Eliminate data silos. Build clean data pipelines that maintain a single source of truth across systems. - Own your services in production: monitoring, alerting, incident response, and post-mortems. AI & Automation: - Build AI/LLM-powered features into internal workflows, automating approvals, knowledge retrieval, reporting, content generation, and operational processes. - Move fast from prototype to production. You know the difference between a demo and a system that works at scale. - Stay current on emerging AI capabilities and proactively identify where they unlock step-change improvements in internal productivity. Collaboration & Influence: - Work directly with business stakeholders to understand pain points and translate them into technical solutions. You don't wait for a spec you help shape it. - Pair with and mentor junior engineers. Raise the technical bar through code reviews, design reviews, and leading by example. - Influence technical direction: propose architectural improvements, challenge assumptions, and drive best practices across the team. Qualifications Senior Engineer Qualifications: - 5+ years of professional software engineering experience, with meaningful time spent building internal tools, platforms, or business systems. - Artificial Intelligence (AI) experience required. - Strong full-stack or backend engineering skills. Proficient in at least one of: Python, Go, TypeScript/Node.js, or Java. - Solid understanding of Cloud Infrastructure (GCP/AWS/Azure), Containerization (Docker/Kubernetes), CI/CD Pipelines, and modern DevOps practices. - Hands-on experience building and maintaining API integrations between third-party SaaS platforms (e.g., Workday, Salesforce, Slack, NetSuite). - Strong data fundamentals: relational databases, data modelling, ETL/ELT pipelines, and working knowledge of SQL. - Comfort with ambiguity. You can take a vague business problem, break it down, and deliver a working solution without heavy handholding. - Clear communicator who can explain technical tradeoffs to non-technical stakeholders. - Experience with workflow orchestration tools (Temporal, Airflow, Prefect) or integration platforms (Workato, Tray.io, MuleSoft) is a plus. - Frontend experience with React, Next.js, or equivalent modern frameworks is a plus. - Familiarity with HRIS, ERP, or people systems data models and processes is a plus. - Experience at a high-growth or AI-native company is a plus. - Contributions to developer experience tooling, CLIs, or internal SDKs is a plus. - Experience building or integrating AI/LLM-powered features not just experimenting, but shipping to real users is a plus. Benefits include medical insurance, Dental, Vision, Savings Plan Options, PTO, etc. Looking to hire a Senior Engineer in San Francisco, CA or in other cities? Our IT recruiting agencies and staffing companies can help. We help companies that are looking to hire Senior Engineers for jobs in San Francisco, California and in other cities too. Please contact our IT recruiting agencies and IT staffing companies today! Additional Information Please check out all of our jobs at .
08/05/2026
Full time
Job Description Job Description Job Description Senior Engineer, Internal Tools, Artificial Intelligence (AI) Required, Work From Home The Senior AI Engineer is the engineering backbone of the internal tools team. Building and maintaining the platforms that every team in the company relies on. Your work directly increases organizational efficiency and enables teams to move faster. The Senior AI Engineer will own systems end-to-end, from scoping and architecture through production deployment and iteration, connecting multiple business systems into a seamless, reliable internal ecosystem. This position is 100% Remote. Senior Engineer Responsibilities: Build & Ship: - Design, build, and maintain internal platforms and tools that serve People, Finance, Ops, Sales, and Engineering teams. - Own features, end-to-end requirements, architecture, implementation, testing, deployment, and monitoring. - Write clean, well-tested, production-grade code. You hold yourself to the same bar as customer-facing products. Architecture & Integration: - Build API-first integrations across the internal ecosystem connecting HRIS, CRM, finance platforms, knowledge management, and developer tools into a coherent stack. - Design for reliability, performance, and scale what you build today must hold as the company grows 5-10x. - Eliminate data silos. Build clean data pipelines that maintain a single source of truth across systems. - Own your services in production: monitoring, alerting, incident response, and post-mortems. AI & Automation: - Build AI/LLM-powered features into internal workflows, automating approvals, knowledge retrieval, reporting, content generation, and operational processes. - Move fast from prototype to production. You know the difference between a demo and a system that works at scale. - Stay current on emerging AI capabilities and proactively identify where they unlock step-change improvements in internal productivity. Collaboration & Influence: - Work directly with business stakeholders to understand pain points and translate them into technical solutions. You don't wait for a spec you help shape it. - Pair with and mentor junior engineers. Raise the technical bar through code reviews, design reviews, and leading by example. - Influence technical direction: propose architectural improvements, challenge assumptions, and drive best practices across the team. Qualifications Senior Engineer Qualifications: - 5+ years of professional software engineering experience, with meaningful time spent building internal tools, platforms, or business systems. - Artificial Intelligence (AI) experience required. - Strong full-stack or backend engineering skills. Proficient in at least one of: Python, Go, TypeScript/Node.js, or Java. - Solid understanding of Cloud Infrastructure (GCP/AWS/Azure), Containerization (Docker/Kubernetes), CI/CD Pipelines, and modern DevOps practices. - Hands-on experience building and maintaining API integrations between third-party SaaS platforms (e.g., Workday, Salesforce, Slack, NetSuite). - Strong data fundamentals: relational databases, data modelling, ETL/ELT pipelines, and working knowledge of SQL. - Comfort with ambiguity. You can take a vague business problem, break it down, and deliver a working solution without heavy handholding. - Clear communicator who can explain technical tradeoffs to non-technical stakeholders. - Experience with workflow orchestration tools (Temporal, Airflow, Prefect) or integration platforms (Workato, Tray.io, MuleSoft) is a plus. - Frontend experience with React, Next.js, or equivalent modern frameworks is a plus. - Familiarity with HRIS, ERP, or people systems data models and processes is a plus. - Experience at a high-growth or AI-native company is a plus. - Contributions to developer experience tooling, CLIs, or internal SDKs is a plus. - Experience building or integrating AI/LLM-powered features not just experimenting, but shipping to real users is a plus. Benefits include medical insurance, Dental, Vision, Savings Plan Options, PTO, etc. Looking to hire a Senior Engineer in San Francisco, CA or in other cities? Our IT recruiting agencies and staffing companies can help. We help companies that are looking to hire Senior Engineers for jobs in San Francisco, California and in other cities too. Please contact our IT recruiting agencies and IT staffing companies today! Additional Information Please check out all of our jobs at .
Senior Site Reliability Engineer- San Francisco, CA, the US
Kody San Francisco, California
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
08/05/2026
Full time
Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will own production observability, incident response, service-level management, and cloud infrastructure reliability across mission-critical payment processing systems operating in Europe, Asia, and North America. Responsibilities Participate in a follow-the-sun production on-call rotation as a primary incident responder. Diagnose, triage, mitigate, and coordinate resolution of production incidents across payment services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure. Define and maintain SLOs, SLIs, error budgets, alerting standards, and operational readiness processes. Drive reliability improvements through automation, observability, capacity planning, performance optimization, and post-incident reviews. Partner with engineering teams to improve resilience, security, and operational maturity in PCI-DSS-regulated environments. Lead incident management during SEV1/SEV2 events and improve response effectiveness and MTTR. Requirements 5+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Cloud Infrastructure roles supporting mission-critical production systems. Strong hands-on experience with AWS, Kubernetes (EKS), Terraform, PostgreSQL, Redis, Kafka, Linux, networking, and modern observability platforms. Deep understanding of distributed systems, cloud-native architectures, high availability, disaster recovery, capacity planning, and performance optimization. Proven experience operating payment, banking, fintech, or other highly regulated systems with stringent security, compliance, and uptime requirements. Strong knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, alert governance, and operational excellence. Leadership & Operational Excellence Demonstrates strong ownership and accountability, taking end-to-end responsibility for service reliability and customer impact. Possesses a strong sense of urgency during production incidents while maintaining sound judgment and structured decision-making under pressure. Applies a systematic and methodical approach to troubleshooting, root-cause analysis, and incident resolution in complex distributed environments. Data-driven mindset with the ability to leverage metrics, telemetry, trends, and service-level indicators to prioritize reliability investments and operational improvements. Continuously drives engineering excellence through iterative improvement, automation, standardization, and elimination of operational toil. Proven ability to lead cross-functional incident response efforts, coordinate stakeholders, and communicate effectively during high-severity production events. Champions a culture of operational readiness, continuous learning, post-incident improvement, and blameless accountability. Demonstrates strong mentoring and technical leadership skills, influencing engineering teams to build reliable, scalable, and resilient systems by design. Benefits Competitive packages aligned with California market standards Lead a dynamic and innovative team in a very rapidly growing company Collaborative, inclusive environment where your contributions are recognized and valued
Senior Platform Engineer
OnHires San Francisco, California
Job Description Job Description About Us: We are building a robust, scalable trading platform to serve high-traffic, latency-sensitive applications. Our infrastructure leverages state-of-the-art technologies to support real-time trading while providing unparalleled reliability and performance. Join us to shape the future of our platform and engineering culture. Job Summary: We are looking for a Senior DevOps & Platform Engineer to lead the design, implementation, and management of our AWS-centric infrastructure. You will play a pivotal role in maximizing the velocity of our product engineering team, ensuring platform scalability, reliability, and security. This is a high-impact role, combining elements of DevOps, Platform Engineering, and Site Reliability Engineering (SRE). You will champion best practices, shape the engineering culture, and ensure our platform is robust, efficient, and ready for the future. Key Responsibilities: Platform Engineering Infrastructure Design: Architect and implement scalable infrastructure to support the deployment and management of our trading platform. Developer Tooling: Build and maintain internal tools to streamline developer workflows, including advanced CI/CD pipelines. Infrastructure as Code (IaC): Champion IaC practices using Terraform, CloudFormation, or Pulumi. Core Services Management: Manage and optimize platform-critical services such as: NATS Cluster RabbitMQ AWS RDS PostgreSQL Redis Cluster DevOps Automation and CI/CD: Automate and optimize deployment processes to ensure seamless continuous integration and delivery. Container Orchestration: Manage and scale containerized workloads using Kubernetes and Docker. Cloud Optimization: Monitor and optimize cloud resource usage for performance and cost efficiency. Site Reliability Engineering (SRE) Reliability Metrics: Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Monitoring & Observability: Implement observability tools and dashboards (e.g., Prometheus, Datadog, Grafana) for real-time system monitoring. Incident Management: Lead incident response efforts, conduct root cause analysis, and implement actionable postmortem reviews. Infrastructure Management AWS Expertise: Architect and manage cloud-based systems to handle high-traffic, latency-sensitive applications. Disaster Recovery: Implement robust disaster recovery and business continuity strategies, including backups and multi-region failover. Security Practices: Collaborate with security teams to enforce best practices for IAM, encryption, and compliance. Collaboration & Leadership Cross-Team Collaboration: Partner with software engineers to design infrastructure solutions tailored to their application needs. Culture Building: Help shape the engineering culture, promoting a philosophy of security, velocity, and reliability. Mentorship: Mentor junior engineers and document best practices to drive knowledge sharing and operational excellence. Long-Term Tech Evolution Backend Transition: Contribute to evolving our backend microservices (currently NodeJS, with some Python and C#) towards Go and Rust. Third-Party Integration: Evaluate and integrate critical third-party software and infrastructure, such as payment gateways and mobility stacks. Your Impact: Simplify infrastructure concerns for product teams to accelerate builds, deployments, and scaling. Advocate for modern practices like Zero Trust Networking and continuously improve platform architecture. Balance the demands of product velocity with a well-managed, secure, and scalable platform. Required Skills & Experience: Technical Expertise Cloud Experience: 5-8+ years of hands-on experience with cloud platforms, particularly AWS, including services like EC2, RDS, S3, Lambda, and VPC. Containerization: Proficiency with Docker and Kubernetes (EKS) or ECS. Infrastructure as Code (IaC): Strong experience with Terraform, CloudFormation, or Pulumi. Programming Skills: Proficiency in at least one programming language (e.g., Python, Go, TypeScript/JavaScript, Ruby, Java). DevOps & SRE CI/CD Pipelines: Expertise in building and maintaining CI/CD workflows using tools like GitLab CI, Jenkins, or GitHub Actions. Monitoring Tools: Experience with observability platforms (e.g., Prometheus, Datadog, Grafana). Incident Management: Proven ability to handle incident response, root cause analysis, and postmortem reviews. Soft Skills Problem-Solving: Ability to research, design, and deliver solutions to complex infrastructure challenges. Collaboration: Experience working directly with product engineers to improve workflows incrementally. Leadership: Ownership mindset with the ability to mentor team members and advocate for best practices. Preferred Skills (Nice-to-Have): Familiarity with backend languages like Go or Rust. AWS certifications (e.g., Solutions Architect, DevOps Engineer). Experience with networking concepts (e.g., load balancers, DNS, VPNs) and traffic optimization. Knowledge of emerging CNCF technologies and CI/CD trends. What We Offer: Competitive salary with future equity options Opportunities to work with cutting-edge technologies and evolve our platform. Flexible working hours and a remote-friendly environment. Professional growth through certifications, conferences, and internal training. Collaborative culture focused on innovation and operational excellence.
08/05/2026
Full time
Job Description Job Description About Us: We are building a robust, scalable trading platform to serve high-traffic, latency-sensitive applications. Our infrastructure leverages state-of-the-art technologies to support real-time trading while providing unparalleled reliability and performance. Join us to shape the future of our platform and engineering culture. Job Summary: We are looking for a Senior DevOps & Platform Engineer to lead the design, implementation, and management of our AWS-centric infrastructure. You will play a pivotal role in maximizing the velocity of our product engineering team, ensuring platform scalability, reliability, and security. This is a high-impact role, combining elements of DevOps, Platform Engineering, and Site Reliability Engineering (SRE). You will champion best practices, shape the engineering culture, and ensure our platform is robust, efficient, and ready for the future. Key Responsibilities: Platform Engineering Infrastructure Design: Architect and implement scalable infrastructure to support the deployment and management of our trading platform. Developer Tooling: Build and maintain internal tools to streamline developer workflows, including advanced CI/CD pipelines. Infrastructure as Code (IaC): Champion IaC practices using Terraform, CloudFormation, or Pulumi. Core Services Management: Manage and optimize platform-critical services such as: NATS Cluster RabbitMQ AWS RDS PostgreSQL Redis Cluster DevOps Automation and CI/CD: Automate and optimize deployment processes to ensure seamless continuous integration and delivery. Container Orchestration: Manage and scale containerized workloads using Kubernetes and Docker. Cloud Optimization: Monitor and optimize cloud resource usage for performance and cost efficiency. Site Reliability Engineering (SRE) Reliability Metrics: Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Monitoring & Observability: Implement observability tools and dashboards (e.g., Prometheus, Datadog, Grafana) for real-time system monitoring. Incident Management: Lead incident response efforts, conduct root cause analysis, and implement actionable postmortem reviews. Infrastructure Management AWS Expertise: Architect and manage cloud-based systems to handle high-traffic, latency-sensitive applications. Disaster Recovery: Implement robust disaster recovery and business continuity strategies, including backups and multi-region failover. Security Practices: Collaborate with security teams to enforce best practices for IAM, encryption, and compliance. Collaboration & Leadership Cross-Team Collaboration: Partner with software engineers to design infrastructure solutions tailored to their application needs. Culture Building: Help shape the engineering culture, promoting a philosophy of security, velocity, and reliability. Mentorship: Mentor junior engineers and document best practices to drive knowledge sharing and operational excellence. Long-Term Tech Evolution Backend Transition: Contribute to evolving our backend microservices (currently NodeJS, with some Python and C#) towards Go and Rust. Third-Party Integration: Evaluate and integrate critical third-party software and infrastructure, such as payment gateways and mobility stacks. Your Impact: Simplify infrastructure concerns for product teams to accelerate builds, deployments, and scaling. Advocate for modern practices like Zero Trust Networking and continuously improve platform architecture. Balance the demands of product velocity with a well-managed, secure, and scalable platform. Required Skills & Experience: Technical Expertise Cloud Experience: 5-8+ years of hands-on experience with cloud platforms, particularly AWS, including services like EC2, RDS, S3, Lambda, and VPC. Containerization: Proficiency with Docker and Kubernetes (EKS) or ECS. Infrastructure as Code (IaC): Strong experience with Terraform, CloudFormation, or Pulumi. Programming Skills: Proficiency in at least one programming language (e.g., Python, Go, TypeScript/JavaScript, Ruby, Java). DevOps & SRE CI/CD Pipelines: Expertise in building and maintaining CI/CD workflows using tools like GitLab CI, Jenkins, or GitHub Actions. Monitoring Tools: Experience with observability platforms (e.g., Prometheus, Datadog, Grafana). Incident Management: Proven ability to handle incident response, root cause analysis, and postmortem reviews. Soft Skills Problem-Solving: Ability to research, design, and deliver solutions to complex infrastructure challenges. Collaboration: Experience working directly with product engineers to improve workflows incrementally. Leadership: Ownership mindset with the ability to mentor team members and advocate for best practices. Preferred Skills (Nice-to-Have): Familiarity with backend languages like Go or Rust. AWS certifications (e.g., Solutions Architect, DevOps Engineer). Experience with networking concepts (e.g., load balancers, DNS, VPNs) and traffic optimization. Knowledge of emerging CNCF technologies and CI/CD trends. What We Offer: Competitive salary with future equity options Opportunities to work with cutting-edge technologies and evolve our platform. Flexible working hours and a remote-friendly environment. Professional growth through certifications, conferences, and internal training. Collaborative culture focused on innovation and operational excellence.
Senior Site Reliability Engineer - Compute Platforms
Five9 San Ramon, California
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer - Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices. This is a deeply technical role requiring expert-level understanding of compute hardware management, Kubernetes, OpenStack, hypervisors and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications. Key Responsibilities Lead the architecture and design of enterprise compute and hypervisor platform solutions across hardware, OS, virtualization, cloud orchestration, and container orchestration layers Define standards and automation frameworks for bare metal provisioning and lifecycle management Design and implement Bare Metal as a Service (BMaaS) capabilities for scalable infrastructure consumption Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD) Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester Design and maintain PXE-based provisioning environments leveraging Redfish APIs for large-scale server deployments Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation. Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback. Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation Evaluate and standardize enterprise hardware platforms to meet performance, scalability, and reliability requirements Produce detailed high-level and low-level design documentation , build guides, and operational handoff materials Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-ready Participate in on-call escalation support for complex platform-related issues Collaborate globally on change management , documentation, and operational best practices Minimum Qualifications 6 + years of experience in infrastructure engineering, platform engineering, or DevOps with a strong focus on Compute system design Proven experience designing and automating bare metal compute environments at scale Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms Practical experience using Redfish APIs for hardware provisioning, power management, and remote lifecycle operations Deep expertise with Ubuntu Linux in enterprise environments Strong Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Experience designing and deploying production-grade Kubernetes clusters Strong background with enterprise compute hardware platforms , including Cisco UCS, Dell PowerEdge, Supermicro systems & HPE Proficiency with Infrastructure as Code tools (e.g., Terraform, Ansible, or similar) Experience building or supporting CI/CD pipelines for infrastructure and platform automation Strong scripting skills in Python, Bash, or similar languages Demonstrated ability to produce clear, structured technical design documentation Excellent written and verbal communication skills Bachelor's degree in computer science or equivalent professional experience Preferred Qualifications OpenStack, Ubuntu KVM administration. BareMetal as a Service (PXE, Redfish). Kubernetes on BareMetal CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. Background in telco, edge cloud, or large enterprise environments. Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) Master's degree in computer science, IT, Engineering, or a related field preferred; equivalent experience and relevant industry certifications will also be considered What You'll Get A collaborative team that's deeply invested in infrastructure excellence. Complex technical challenges that require creative, scalable solutions. The opportunity to shape a next-generation private cloud platform-built reliability Access to the latest tools, frameworks, and upstream project developments Skills and Attributes: Analytical Thinking & Problem Solving: Demonstrated ability to translate complex, cross-domain requirements into scalable and resilient cloud infrastructure and automation solutions Collaboration & Teamwork: Strong interpersonal and communication skills with a proven track record of effective collaboration across multidisciplinary teams, including developers, operations, security, and product stakeholders Mentorship & Leadership: Passionate about knowledge-sharing and mentorship, with experience guiding junior engineers and fostering a team culture of continuous learning, innovation, and technical excellence in cloud engineering and DevOps practices Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally, the total compensation package for this position may also include an annual performance bonus, stock, and/or other applicable incentive compensation plans. Our total reward package also includes: Health, dental, and vision coverage, beginning on the first day of employment. Five9 covers 100% of the employee portion of the health, dental and vision coverage and shares a high portion of the dependent cost. We also offer Short & Long-Term Disability, Basic Life Insurance, and a 401k saving plan with employer matching. Access to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for all covered employees and their covered dependents. Generous employee stock purchase plan. Paid Time Off, Company paid holidays, paid volunteer hours and 12 weeks paid parental leave. All compensation and benefits are subject to the requirements and restrictions set forth in the applicable plan documents and any written agreements between the parties. The US base salary range for this role is below. $82,300-$228,800 USD Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better we are. Five9 is an equal opportunity employer. View our privacy policy, including our privacy notice to California residents here: -pt/legal. Note: Five9 will never request that an applicant send money as a prerequisite for commencing employment with Five9.
08/05/2026
Full time
Job Description Job Description Join us in bringing joy to customer experience. Five9 is a leading provider of cloud contact center software, bringing the power of cloud innovation to customers worldwide. Living our values everyday results in our team-first culture and enables us to innovate, grow, and thrive while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer - Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices. This is a deeply technical role requiring expert-level understanding of compute hardware management, Kubernetes, OpenStack, hypervisors and extensive working knowledge on Linux Operating systems. You will also collaborate with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated services that serve high-throughput, mission-critical applications. Key Responsibilities Lead the architecture and design of enterprise compute and hypervisor platform solutions across hardware, OS, virtualization, cloud orchestration, and container orchestration layers Define standards and automation frameworks for bare metal provisioning and lifecycle management Design and implement Bare Metal as a Service (BMaaS) capabilities for scalable infrastructure consumption Architect and design Kubernetes platforms on bare metal with QoS and Affinity (ArgoCD) Architect and validate automated deployments of operating systems and hypervisors including Ubuntu and Harvester Design and maintain PXE-based provisioning environments leveraging Redfish APIs for large-scale server deployments Develop Infrastructure-as-Code using Ansible, Terraform, Helm and Git, with Python/Bash automation. Implement CI/CD pipelines for infrastructure updates, patching, upgrades, testing, and rollback. Design automated workflows for server build, firmware lifecycle management, patching, and hardware validation Evaluate and standardize enterprise hardware platforms to meet performance, scalability, and reliability requirements Produce detailed high-level and low-level design documentation , build guides, and operational handoff materials Perform deep troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network, storage, and platform teams to ensure designs are supportable and production-ready Participate in on-call escalation support for complex platform-related issues Collaborate globally on change management , documentation, and operational best practices Minimum Qualifications 6 + years of experience in infrastructure engineering, platform engineering, or DevOps with a strong focus on Compute system design Proven experience designing and automating bare metal compute environments at scale Strong hands-on experience with PXE boot, network-based OS provisioning, and automated server imaging Experience implementing or supporting Bare Metal as a Service (BMaaS) platforms Practical experience using Redfish APIs for hardware provisioning, power management, and remote lifecycle operations Deep expertise with Ubuntu Linux in enterprise environments Strong Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Experience designing and deploying production-grade Kubernetes clusters Strong background with enterprise compute hardware platforms , including Cisco UCS, Dell PowerEdge, Supermicro systems & HPE Proficiency with Infrastructure as Code tools (e.g., Terraform, Ansible, or similar) Experience building or supporting CI/CD pipelines for infrastructure and platform automation Strong scripting skills in Python, Bash, or similar languages Demonstrated ability to produce clear, structured technical design documentation Excellent written and verbal communication skills Bachelor's degree in computer science or equivalent professional experience Preferred Qualifications OpenStack, Ubuntu KVM administration. BareMetal as a Service (PXE, Redfish). Kubernetes on BareMetal CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. Background in telco, edge cloud, or large enterprise environments. Ubuntu Certifications, CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) Master's degree in computer science, IT, Engineering, or a related field preferred; equivalent experience and relevant industry certifications will also be considered What You'll Get A collaborative team that's deeply invested in infrastructure excellence. Complex technical challenges that require creative, scalable solutions. The opportunity to shape a next-generation private cloud platform-built reliability Access to the latest tools, frameworks, and upstream project developments Skills and Attributes: Analytical Thinking & Problem Solving: Demonstrated ability to translate complex, cross-domain requirements into scalable and resilient cloud infrastructure and automation solutions Collaboration & Teamwork: Strong interpersonal and communication skills with a proven track record of effective collaboration across multidisciplinary teams, including developers, operations, security, and product stakeholders Mentorship & Leadership: Passionate about knowledge-sharing and mentorship, with experience guiding junior engineers and fostering a team culture of continuous learning, innovation, and technical excellence in cloud engineering and DevOps practices Work Location: This role is fully remote for candidates who reside outside the 30 mile radius of one of our offices. For candidates who reside within a 30 mile radius of one of our offices, this role is Hybrid and would require 3 days a week (T, W, TH) in office. As part of our continued commitment to diversity, equity, and inclusion, Five9 supports pay transparency during the entire recruitment process. Actual compensation packages are based on several factors that are unique to each candidate including, but not limited to: skill set, depth of experience, certifications, and specific work location. The range displayed reflects the minimum and maximum target for new hire salaries for the job across the United States. Your recruiter can share more about the specific compensation package during your hiring process. Additionally, the total compensation package for this position may also include an annual performance bonus, stock, and/or other applicable incentive compensation plans. Our total reward package also includes: Health, dental, and vision coverage, beginning on the first day of employment. Five9 covers 100% of the employee portion of the health, dental and vision coverage and shares a high portion of the dependent cost. We also offer Short & Long-Term Disability, Basic Life Insurance, and a 401k saving plan with employer matching. Access to an innovative mental health support platform that offers personalized care and resources in areas such as: therapy, coaching and self-guided mindfulness exercises for all covered employees and their covered dependents. Generous employee stock purchase plan. Paid Time Off, Company paid holidays, paid volunteer hours and 12 weeks paid parental leave. All compensation and benefits are subject to the requirements and restrictions set forth in the applicable plan documents and any written agreements between the parties. The US base salary range for this role is below. $82,300-$228,800 USD Five9 embraces diversity and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better we are. Five9 is an equal opportunity employer. View our privacy policy, including our privacy notice to California residents here: -pt/legal. Note: Five9 will never request that an applicant send money as a prerequisite for commencing employment with Five9.
Senior Cloud Engineer
SmithRx San Jose, California
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
08/05/2026
Full time
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
Sr. Staff Site Reliability Engineer
Obsidian Security Palo Alto, California
Job Description Job Description Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast Asia, Australia, and New Zealand, including many of the world's largest Fortune 1000 and Global 2000 companies. Founded in 2017 and backed by top investors like Greylock, Obsidian was built to close a critical gap: securing SaaS apps where business happens-Microsoft 365, Salesforce, and hundreds more. The company does this by offering a complete SaaS security platform to reduce risk, detect and respond to threats, and prevent breaches at the source. Obsidian was built by leaders who redefined endpoint and identity security at CrowdStrike, Okta, Cylance, and Carbon Black. Now, they're transforming how SaaS is secured. With AI driving rapid SaaS growth and complexity, agentic AI tools gain privileged access to sensitive data through integrations, creating new risks most security tools miss. Obsidian uniquely detects anomalous OAuth token activity and manages integration risks. Major announcements are on the horizon. Recognizing that SaaS security needs to evolve, Obsidian enables growing organizations to start with a lightweight, prevention-focused browser extension and expand coverage over time. With global momentum, a growing partner ecosystem including SentinelOne, Databricks, and Google Cloud, and a major fundraise ahead, Obsidian is scaling rapidly toward long-term growth and IPO readiness. Sr. Staff Site Reliability Engineer As a Sr. Staff SRE at Obsidian , you will define and drive the company-wide reliability vision for a complex, multi-tenant SaaS platform serving enterprise and financial customers. You will operate as a strategic partner to DevOps and Platform Engineering leadership, shaping a unified reliability strategy that scales across the organization. Your core mandate: ensure Obsidian detects, diagnoses, and communicates system issues before customers are impacted-consistently and predictably. This is a hands-on technical role that involves architecting and leading the implementation of systems that handle real-world complexity, including upstream SaaS dependencies, sparse and noisy signals, and mission-critical enterprise workloads. Key Responsibilities Reliability Strategy & Architecture - Define and lead long-term reliability strategy across services. Establish end-to-end system visibility frameworks and guide architecture for observability, detection, and resilience. Cross-Org Leadership - Partner across teams to embed reliability, standardize SLI/SLOs, and serve as a technical escalation expert. Detection & Observability - Build intelligent detection systems (anomaly detection, connector health models) and enable self-service observability. Incident Management - Define and evolve a tiered incident communication strategy , improve response practices, and lead postmortems to strengthen reliability and customer trust. Execution - Contribute hands-on to system design, monitoring, and debugging across distributed systems and data pipelines. Required Qualifications 5+ years in SRE, Production Engineering, or related roles 3+ years operating at a senior or technical leadership level (Staff or equivalent scope) Deep expertise in: AWS and/or GCP Kubernetes and Helm Observability stacks (Prometheus, Grafana, or equivalent) CI/CD systems (GitLab CI/CD, ArgoCD, etc.) Proven experience designing and scaling reliability systems for multi-tenant SaaS platforms Strong debugging and systems thinking across distributed microservices and legacy systems Demonstrated ability to lead initiatives that improve incident detection, response, and system resilience Hands-on engineering approach with a track record of building-not just configuring-reliability systems Preferred Qualifications Experience in B2B SaaS serving enterprise or financial customers Familiarity with third-party SaaS connector architectures and ingestion patterns Experience building anomaly detection or intelligent alerting systems Experience designing customer-facing status pages and incident communication frameworks Why This Role Drive org-wide reliability strategy Own and build new detection & observability systems Tackle complex distributed systems challenges Safeguard critical infrastructure for financial customers What Success Looks Like Issues caught and resolved before customer impact Reliability is measurable and continuously improving Teams self-serve observability with scalable tools Clear, proactive incident communication builds trust Reliability becomes a competitive advantage Employee Benefits Our competitive benefits packages are designed to support our employees' well-being, both at work and at home. Our US based employees enjoy: Competitive compensation with equity and 401k Comprehensive healthcare with dental and vision coverage Flexible paid time off and paid holiday time off 12 weeks of new parent or family leave Personal and professional development resources For more details on our US benefits, or for information on our international benefits, please see here. Pay Transparancy Please note that the base pay range is a guideline and for candidates who receive an offer, the base pay will vary based on factors such as work location, as well as the knowledge, skills and experience of the candidate. In addition to a competitive base salary, this position is eligible for equity awards and may be eligible for sales commission or incentive compensation based on the role or function within the company. At Obsidian, we are proud to be an equal-opportunity employer. We value diversity and hire for talent, passion, and compassion. In compliance with federal law, all persons hired will be required to submit satisfactory proof of identity and legal authorization. If you have a need that requires accommodation, please contact Information collected and processed as part of any job applications you choose to submit is subject to Obsidian's Applicant Privacy Policy. Base Salary Range $232,000-$263,000 USD
08/05/2026
Full time
Job Description Job Description Obsidian Security is the leading SaaS security platform, trusted by global enterprises like Snowflake, T-Mobile, and Algolia. We protect 200+ organizations across North America, Europe, the Middle East, Southeast Asia, Australia, and New Zealand, including many of the world's largest Fortune 1000 and Global 2000 companies. Founded in 2017 and backed by top investors like Greylock, Obsidian was built to close a critical gap: securing SaaS apps where business happens-Microsoft 365, Salesforce, and hundreds more. The company does this by offering a complete SaaS security platform to reduce risk, detect and respond to threats, and prevent breaches at the source. Obsidian was built by leaders who redefined endpoint and identity security at CrowdStrike, Okta, Cylance, and Carbon Black. Now, they're transforming how SaaS is secured. With AI driving rapid SaaS growth and complexity, agentic AI tools gain privileged access to sensitive data through integrations, creating new risks most security tools miss. Obsidian uniquely detects anomalous OAuth token activity and manages integration risks. Major announcements are on the horizon. Recognizing that SaaS security needs to evolve, Obsidian enables growing organizations to start with a lightweight, prevention-focused browser extension and expand coverage over time. With global momentum, a growing partner ecosystem including SentinelOne, Databricks, and Google Cloud, and a major fundraise ahead, Obsidian is scaling rapidly toward long-term growth and IPO readiness. Sr. Staff Site Reliability Engineer As a Sr. Staff SRE at Obsidian , you will define and drive the company-wide reliability vision for a complex, multi-tenant SaaS platform serving enterprise and financial customers. You will operate as a strategic partner to DevOps and Platform Engineering leadership, shaping a unified reliability strategy that scales across the organization. Your core mandate: ensure Obsidian detects, diagnoses, and communicates system issues before customers are impacted-consistently and predictably. This is a hands-on technical role that involves architecting and leading the implementation of systems that handle real-world complexity, including upstream SaaS dependencies, sparse and noisy signals, and mission-critical enterprise workloads. Key Responsibilities Reliability Strategy & Architecture - Define and lead long-term reliability strategy across services. Establish end-to-end system visibility frameworks and guide architecture for observability, detection, and resilience. Cross-Org Leadership - Partner across teams to embed reliability, standardize SLI/SLOs, and serve as a technical escalation expert. Detection & Observability - Build intelligent detection systems (anomaly detection, connector health models) and enable self-service observability. Incident Management - Define and evolve a tiered incident communication strategy , improve response practices, and lead postmortems to strengthen reliability and customer trust. Execution - Contribute hands-on to system design, monitoring, and debugging across distributed systems and data pipelines. Required Qualifications 5+ years in SRE, Production Engineering, or related roles 3+ years operating at a senior or technical leadership level (Staff or equivalent scope) Deep expertise in: AWS and/or GCP Kubernetes and Helm Observability stacks (Prometheus, Grafana, or equivalent) CI/CD systems (GitLab CI/CD, ArgoCD, etc.) Proven experience designing and scaling reliability systems for multi-tenant SaaS platforms Strong debugging and systems thinking across distributed microservices and legacy systems Demonstrated ability to lead initiatives that improve incident detection, response, and system resilience Hands-on engineering approach with a track record of building-not just configuring-reliability systems Preferred Qualifications Experience in B2B SaaS serving enterprise or financial customers Familiarity with third-party SaaS connector architectures and ingestion patterns Experience building anomaly detection or intelligent alerting systems Experience designing customer-facing status pages and incident communication frameworks Why This Role Drive org-wide reliability strategy Own and build new detection & observability systems Tackle complex distributed systems challenges Safeguard critical infrastructure for financial customers What Success Looks Like Issues caught and resolved before customer impact Reliability is measurable and continuously improving Teams self-serve observability with scalable tools Clear, proactive incident communication builds trust Reliability becomes a competitive advantage Employee Benefits Our competitive benefits packages are designed to support our employees' well-being, both at work and at home. Our US based employees enjoy: Competitive compensation with equity and 401k Comprehensive healthcare with dental and vision coverage Flexible paid time off and paid holiday time off 12 weeks of new parent or family leave Personal and professional development resources For more details on our US benefits, or for information on our international benefits, please see here. Pay Transparancy Please note that the base pay range is a guideline and for candidates who receive an offer, the base pay will vary based on factors such as work location, as well as the knowledge, skills and experience of the candidate. In addition to a competitive base salary, this position is eligible for equity awards and may be eligible for sales commission or incentive compensation based on the role or function within the company. At Obsidian, we are proud to be an equal-opportunity employer. We value diversity and hire for talent, passion, and compassion. In compliance with federal law, all persons hired will be required to submit satisfactory proof of identity and legal authorization. If you have a need that requires accommodation, please contact Information collected and processed as part of any job applications you choose to submit is subject to Obsidian's Applicant Privacy Policy. Base Salary Range $232,000-$263,000 USD
Senior Cloud Engineer
SmithRx San Francisco, California
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
08/05/2026
Full time
Job Description Job Description Who We Are: SmithRx is a rapidly growing, venture-backed Health-Tech company. Our mission is to disrupt the expensive and inefficient Pharmacy Benefit Management (PBM) sector by building a next-generation drug acquisition platform driven by cutting edge technology, innovative cost saving tools, and best-in-class customer service. With hundreds of thousands of members onboarded since 2016, SmithRx has a solution that is resonating with clients all across the country. We pride ourselves for our mission-driven and collaborative culture that inspires our employees to do their best work. We believe that the U.S healthcare system is in need of transformation, and we come to work each day dedicated to making that change a reality. At our core, we are guided by our company values: Integrity: Our purpose guides our actions and gives us confidence in the path ahead. With unwavering honesty and dependability, we embrace the pressure of challenging the old and exemplify ethical leadership to create the new. Courage: We face continuous challenges with grit and resilience. We embrace the discomfort of the unknown by balancing autonomy with empathy, and ownership with vulnerability. We boldly challenge the status quo to keep moving forward-always. Together: The success of SmithRx reflects the strength of our partnerships and the commitment of our team. Our shared values bind us together and make us one. When one falls, we all fall; when one rises, we all rise. Job Summary: We are looking for an Sr. Cloud Engineer who has hands-on experience building and managing a cloud-based infrastructure. Additionally, this engineer will be responsible for development cycles in integration/continuous deployment mode, process monitoring, and more broadly, constructing a "safety culture" within the SmithRx's DevSecOps practice. Our user base is currently doubling annually, and you would share the responsibility of orchestrating a reliable, sustainable, and scalable infrastructure. What you will do: Help build and maintain a container based infrastructure that is elegant, redundant, scalable and compliant, and support the rest of the team doing the same. Be part of SmithRx Agile development team to deliver an end-to-end automation of deployment, monitoring, and infrastructure management in AWS. . Gain a deep understanding of the challenges that SmithRx faces, technical and otherwise; collaborate with other teams to identify and carry out effective solutions. Work closely with our development team to develop and maintain CI/CD pipelines in a reproducible and secure manner. Monitor and troubleshoot infrastructure issues, and perform root cause analysis when necessary. Collaborate with developers to ensure that applications and services are built with scalability, reliability, and security in mind. Organize the highest levels of systems and infrastructure availability, acting proactively Be a pillar of a collaborative learning culture through exploration of new technologies, application of best practices, and any other innovations you would like to experiment with. Develop custom scripts to increase system efficiency and lower the human intervention time on any tasks Be effective in maintaining SmithRX security program controls and best practices. Understand the health regulatory space and maintain continuous compliance on frameworks like HIPAA, and SOC2. Make pragmatic decisions about technical tradeoffs, infrastructure costs, and resource utilization. Be a part of on-call PagerDuty rotations. What you will bring to SmithRx: 5+ years of experience in Cloud Engineering. BS or advanced degree in computer science or other related field. Extensive experience working in containerized Cloud Native environments, specifically AWS and Kubernetes/EKS. Experience using modern monitoring tools like Cloudwatch, Event Bridge, DataDog etc., and establishing metrics, monitoring, alarming and dashboards. Experience managing change management practices, policies and procedures. Experience deploying and monitoring applications in AWS at scale. Security first mindset, including demonstrated experience building secure development and test environments integrated to CI/CD pipelines and software release cycles. Experience building and maintaining a container based infrastructure and Kubernetes Experience with Infrastructure as Code (Terraform experience a plus), DevOps, SRE concepts and best practices. Experience with infrastructure automation, systems reliability, load balancing, monitoring, logging. Experience with FinOps practices and establishing related governance programs. Experience with fully automating CI/CD pipelines with associated tools such as GitHub Actions. Experience working in and architecting for regulated environments with data privacy regulations like GDPR, HIPAA preferred. Experience working and managing SQL and NoSQL databases like RDS, Redis, Redshift, DynamoDB PostgreSQL, BigQuery, and Snowflake Strong scripting skills in Python, Shell etc. What SmithRx Offers You: Highly competitive wellness benefits including Medical, Pharmacy, Dental, Vision, and Life Insurance and AD&D Insurance Flexible Spending Benefits 401(k) Retirement Savings Program Short-term and long-term disability Discretionary Paid Time Off Paid Company Holidays Wellness Benefits Commuter Benefits Paid Parental Leave benefits Employee Assistance Program (EAP) Well-stocked kitchen in office locations Professional development and training opportunities Location: U.S. Remote Compensation & Benefits Base Salary: The range listed above reflects our standard pay scale and actual pay will vary based on work location, job level, job-related knowledge, skills, and experience. Total Rewards: In addition to base pay, this role may be eligible for bonuses, commissions, or equity. SmithRx offers a variety of benefits to help you live well, including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance and 401(k). Individual offers are tailored to your geography, experience, skills, and education. Your recruiter can share more specific details during the hiring process. In the meantime, feel free to explore our comprehensive benefits here.
Lead Software Engineer
Visa Technology and Operations LLC Foster City, California
About Us Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid. At Visa, you'll have the opportunity to create impact at scale - tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world. Join Visa and do work that matters - to you, to your community, and to the world. Progress starts with you. Job Description We are seeking a highly experienced Lead Software Engineer to join Visa's Shared Services and Cloud Product Development organization. This role is ideal for a senior technologist who thrives as a lead generalist, operating across infrastructure, security, scalability, and platform engineering, while building foundational services that enable Visa's global product ecosystem. In this role, you will design and deliver Java/J2EE-based shared services that power multiple Visa products and platforms, influencing internal engineering standards for scalability, security, reliability, and reusability. You will collaborate across Product, Engineering, Security, SRE, and Architecture teams to build resilient, cloud-native systems that support Visa's mission at global scale. The Work Itself Design, build, and evolve core shared services that support Visa's product development ecosystem and reach approximately 40% of the world's population. Lead the development of scalable, secure, and reusable Java/J2EE-based services that serve as foundational platforms for multiple product teams. Act as a lead generalist, contributing across application architecture, infrastructure, security, and scalability to ensure shared services are enterprise-grade and future-ready. Collaborate cross-functionally to create architecture and design artifacts and deliver best-in-class software solutions used across Visa's technical offerings. Drive continuous improvements in product quality, platform reliability, and operational excellence across shared services. Develop robust, highly available systems serving diverse use cases, including consumer-facing products, B2B platforms, and business-to-government solutions. Leverage modern technologies to build the next generation of Payment Services, Transaction Platforms, Real-Time Payments, and Buy Now Pay Later capabilities. Contribute to team and organizational growth through mentorship, technical leadership, and continuous learning initiatives. Essential Functions Provide deep technical leadership within Shared Services Product Development, with a strong understanding of how shared platforms enable downstream product innovation. Lead discovery and design sessions with product partners to translate business requirements into scalable, secure shared-service architectures. Define and formalize best practices and standards for Java/J2EE development, API design, security, and service scalability. Lead planning, piloting, and integration of new platform capabilities, including infrastructure enhancements, security controls, and AI-enabled tooling. Design and implement cloud-native, containerized services using Docker and Kubernetes, ensuring high availability, fault tolerance, and elastic scalability. Partner closely with Security teams to implement secure API gateways, authentication and authorization mechanisms, and compliance-driven architectures. Analyze systemic patterns across defects, incidents, and performance metrics, driving long-term platform improvements. This is a hybrid position. Expectation of days in the office will be confirmed by your Hiring Manager. Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager. Qualifications Basic Qualifications • 10+ years of relevant work experience with a Bachelor's Degree or at least 7 years of work experience with an Advanced degree (e.g. Masters, MBA, JD, MD) or 4 years of work experience with a PhD, OR 13+ years of relevant work experience. Preferred Qualifications • 12 or more years of work experience with a Bachelor's Degree or 8-10 years of experience with an Advanced Degree (e.g. Masters, MBA, JD, MD) or 6+ years of work experience with a PhD • 13+ years of relevant work experience in lieu of a degree. • 8-10 years of experience with an Advanced Degree, or 6+ years of work experience with a PhD. • Proven ability to define technical needs, develop execution plans, coordinate resources, and deliver results across multiple initiatives. • Demonstrated success leading multiple projects simultaneously and resolving priority or scheduling conflicts. • Strong understanding of container-based cloud architectures, including Kubernetes and Docker. • Experience with cloud migration and multi-cloud strategies. • Experience designing elastic, scalable architectural patterns for high-traffic web applications. • Solid understanding of Service and IT Operations Management and DevOps operating models, including deployment and capacity planning. • Strong experience with enterprise integration, including RESTful web services and API-first architectures. • Deep understanding of security requirements, industry standards, and modern security threats and mitigation strategies. • Experience designing and implementing secure API gateway solutions with dynamic security standards. • Experience with CI/CD-driven architectures and automation-first development practices. • Experience developing commercial software on Unix/Linux platforms. • Proven track record in a technical leadership role, including ownership of system design and delivery. • Experience with consumer-facing application development at scale. • Strong verbal, written, and interpersonal communication skills with both technical and non-technical stakeholders. • Excellent collaboration skills and the ability to influence cross-functional teams without direct authority. • Demonstrated passion for mentoring, team building, and continuous improvement. The Skills You Bring • Technical Leadership: Proven experience leading complex software initiatives across distributed systems and shared service platforms. • Primary Language Expertise: Strong expertise in Java (Core Java, J2EE, RESTful services). • Architecture & Scalability: Experience designing highly available, scalable, and secure distributed systems. • Cloud & Infrastructure: Hands-on experience with container-based architectures using Docker and Kubernetes, and cloud platforms (AWS, Azure, or GCP). • Security: Strong understanding of secure API design, authentication/authorization, secure gateway patterns, and enterprise security standards. • DevOps & Operations: Experience with CI/CD, deployment automation, observability, capacity planning, and production operations. • AI Enablement: Familiarity with integrating AI tools and services to improve platform intelligence and developer productivity. • Collaboration: Strong partnership skills working with Product, QA, DevOps, SRE, and Agile/Scrum teams. • Mentorship: Experience coaching engineers on technical excellence and career development. U.S. Applicants Only The estimated salary range for this position is $192,300.00 to $ 307,600.00 USD per year, which may include potential sales incentive payments (if applicable). Salary may vary depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position may be eligible for bonus and equity.Visa has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program. Work Hours Varies upon the needs of the department. Travel Requirements This position requires travel 5-10% of the time. Mental/Physical Requirements This position will be performed in an office setting. The position will require the incumbent to sit and stand at a desk, communicate in person and by telephone, frequently operate standard office equipment, such as telephones and computers. Visa is an EEO Employer Qualified applicants will receive consideration for employment without regard to race, color religion, sex, national origin, sexual orientation, gender identity, disability or protect veteran status. Visa will also consider for employment qualified applicants with criminal histories in a manner consistent with the EEOC guidelines and applicable local law.
08/05/2026
Full time
About Us Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid. At Visa, you'll have the opportunity to create impact at scale - tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world. Join Visa and do work that matters - to you, to your community, and to the world. Progress starts with you. Job Description We are seeking a highly experienced Lead Software Engineer to join Visa's Shared Services and Cloud Product Development organization. This role is ideal for a senior technologist who thrives as a lead generalist, operating across infrastructure, security, scalability, and platform engineering, while building foundational services that enable Visa's global product ecosystem. In this role, you will design and deliver Java/J2EE-based shared services that power multiple Visa products and platforms, influencing internal engineering standards for scalability, security, reliability, and reusability. You will collaborate across Product, Engineering, Security, SRE, and Architecture teams to build resilient, cloud-native systems that support Visa's mission at global scale. The Work Itself Design, build, and evolve core shared services that support Visa's product development ecosystem and reach approximately 40% of the world's population. Lead the development of scalable, secure, and reusable Java/J2EE-based services that serve as foundational platforms for multiple product teams. Act as a lead generalist, contributing across application architecture, infrastructure, security, and scalability to ensure shared services are enterprise-grade and future-ready. Collaborate cross-functionally to create architecture and design artifacts and deliver best-in-class software solutions used across Visa's technical offerings. Drive continuous improvements in product quality, platform reliability, and operational excellence across shared services. Develop robust, highly available systems serving diverse use cases, including consumer-facing products, B2B platforms, and business-to-government solutions. Leverage modern technologies to build the next generation of Payment Services, Transaction Platforms, Real-Time Payments, and Buy Now Pay Later capabilities. Contribute to team and organizational growth through mentorship, technical leadership, and continuous learning initiatives. Essential Functions Provide deep technical leadership within Shared Services Product Development, with a strong understanding of how shared platforms enable downstream product innovation. Lead discovery and design sessions with product partners to translate business requirements into scalable, secure shared-service architectures. Define and formalize best practices and standards for Java/J2EE development, API design, security, and service scalability. Lead planning, piloting, and integration of new platform capabilities, including infrastructure enhancements, security controls, and AI-enabled tooling. Design and implement cloud-native, containerized services using Docker and Kubernetes, ensuring high availability, fault tolerance, and elastic scalability. Partner closely with Security teams to implement secure API gateways, authentication and authorization mechanisms, and compliance-driven architectures. Analyze systemic patterns across defects, incidents, and performance metrics, driving long-term platform improvements. This is a hybrid position. Expectation of days in the office will be confirmed by your Hiring Manager. Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager. Qualifications Basic Qualifications • 10+ years of relevant work experience with a Bachelor's Degree or at least 7 years of work experience with an Advanced degree (e.g. Masters, MBA, JD, MD) or 4 years of work experience with a PhD, OR 13+ years of relevant work experience. Preferred Qualifications • 12 or more years of work experience with a Bachelor's Degree or 8-10 years of experience with an Advanced Degree (e.g. Masters, MBA, JD, MD) or 6+ years of work experience with a PhD • 13+ years of relevant work experience in lieu of a degree. • 8-10 years of experience with an Advanced Degree, or 6+ years of work experience with a PhD. • Proven ability to define technical needs, develop execution plans, coordinate resources, and deliver results across multiple initiatives. • Demonstrated success leading multiple projects simultaneously and resolving priority or scheduling conflicts. • Strong understanding of container-based cloud architectures, including Kubernetes and Docker. • Experience with cloud migration and multi-cloud strategies. • Experience designing elastic, scalable architectural patterns for high-traffic web applications. • Solid understanding of Service and IT Operations Management and DevOps operating models, including deployment and capacity planning. • Strong experience with enterprise integration, including RESTful web services and API-first architectures. • Deep understanding of security requirements, industry standards, and modern security threats and mitigation strategies. • Experience designing and implementing secure API gateway solutions with dynamic security standards. • Experience with CI/CD-driven architectures and automation-first development practices. • Experience developing commercial software on Unix/Linux platforms. • Proven track record in a technical leadership role, including ownership of system design and delivery. • Experience with consumer-facing application development at scale. • Strong verbal, written, and interpersonal communication skills with both technical and non-technical stakeholders. • Excellent collaboration skills and the ability to influence cross-functional teams without direct authority. • Demonstrated passion for mentoring, team building, and continuous improvement. The Skills You Bring • Technical Leadership: Proven experience leading complex software initiatives across distributed systems and shared service platforms. • Primary Language Expertise: Strong expertise in Java (Core Java, J2EE, RESTful services). • Architecture & Scalability: Experience designing highly available, scalable, and secure distributed systems. • Cloud & Infrastructure: Hands-on experience with container-based architectures using Docker and Kubernetes, and cloud platforms (AWS, Azure, or GCP). • Security: Strong understanding of secure API design, authentication/authorization, secure gateway patterns, and enterprise security standards. • DevOps & Operations: Experience with CI/CD, deployment automation, observability, capacity planning, and production operations. • AI Enablement: Familiarity with integrating AI tools and services to improve platform intelligence and developer productivity. • Collaboration: Strong partnership skills working with Product, QA, DevOps, SRE, and Agile/Scrum teams. • Mentorship: Experience coaching engineers on technical excellence and career development. U.S. Applicants Only The estimated salary range for this position is $192,300.00 to $ 307,600.00 USD per year, which may include potential sales incentive payments (if applicable). Salary may vary depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position may be eligible for bonus and equity.Visa has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program. Work Hours Varies upon the needs of the department. Travel Requirements This position requires travel 5-10% of the time. Mental/Physical Requirements This position will be performed in an office setting. The position will require the incumbent to sit and stand at a desk, communicate in person and by telephone, frequently operate standard office equipment, such as telephones and computers. Visa is an EEO Employer Qualified applicants will receive consideration for employment without regard to race, color religion, sex, national origin, sexual orientation, gender identity, disability or protect veteran status. Visa will also consider for employment qualified applicants with criminal histories in a manner consistent with the EEOC guidelines and applicable local law.
Sr Software Engineer
Disney Entertainment and ESPN Product & Technology Santa Monica, California
Disney Entertainment and ESPN Product & Technology Technology is at the heart of Disney's past, present, and future. Disney Entertainment and ESPN Product & Technology is a global organization of engineers, product developers, designers, technologists, data scientists, and more - all working to build and advance the technological backbone for Disney's media business globally. The team marries technology with creativity to build world-class products, enhance storytelling, and drive velocity, innovation, and scalability for our businesses. We are Storytellers and Innovators. Creators and Builders. Entertainers and Engineers. We work with every part of The Walt Disney Company's media portfolio to advance the technological foundation and consumer media touch points serving millions of people around the world. Here are a few reasons why we think you'd love working here: Building the future of Disney's media: Our Technologists are designing and building the products and platforms that will power our media, advertising, and distribution businesses for years to come. Reach, Scale & Impact: More than ever, Disney's technology and products serve as a signature doorway for fans' connections with the company's brands and stories. Disney+. Hulu. ESPN. ABC. ABC News and many more. These products and brands - and the unmatched stories, storytellers, and events they carry - matter to millions of people globally. Innovation: We develop and implement groundbreaking products and techniques that shape industry norms, and solve complex and distinctive technical problems. Commerce, Data & Identity provides the core product management functions for areas crucial to Disney's media businesses. These include initiatives and products that power digital commerce, identity, and growth, as well as those that reach uniquely across The Walt Disney Company enterprise, such as messaging and privacy, among others. Additionally, it is responsible for the data engineering, science, and products for Disney Entertainment & ESPN, along with their interconnection with other parts of The Walt Disney Company. The Subscriptions team is looking to expand our team in the NYC office. We are a dynamic team of engineers that support all your Disney, ESPN and Hulu subscriptions at scale! We solve problems at internet-scale that only a handful of teams ever reach. We are seeking a highly skilled and motivated Senior Software Engineer to join our dynamic engineering team. In this role, you will be responsible for designing, developing, and deploying scalable and efficient software solutions using Functional Programming Style Scala and other backend languages. You will work with cutting-edge technologies in cloud computing, containerization, distributed systems, and observability. The ideal candidate should have a strong background in full-stack development, modern DevOps practices, and experience working with cloud-native technologies such as AWS and Kubernetes. Responsibilities Development: Design, develop, and maintain scalable, secure, and efficient software applications using Scala and other backend languages. Cloud Infrastructure: Utilize AWS services (CFN, EC2, Lambda, S3, Dynamo, etc.) to deploy and manage applications in the cloud. Containerization & Orchestration: Implement containerized solutions using Docker, deploy and manage services on Kubernetes. Distributed Systems: Design and build distributed systems that are fault-tolerant, highly available, and scalable. Understand concepts such as event-driven architecture, microservices, and data consistency. Observability & Monitoring: Implement and maintain observability best practices, including tagging, metrics, and logging to provide comprehensive visibility into system performance. Use tools like Datadog to monitor the health and performance of applications in real-time. Collaboration: Work closely with cross-functional teams including product managers, UI/UX designers, and DevOps engineers to define requirements, design systems, and deliver features. Code Quality & Best Practices: Write clean, maintainable, and testable code while following best practices for software development. Perform code reviews and provide mentorship to junior engineers on and offshore. Problem-Solving & Innovation: Tackle complex technical challenges and continuously seek opportunities to improve system performance, scalability, and reliability.AI Best Practices: Follow team's best practices regarding the safe use of LLMs and suggest improvements as appropriate. Basic Qualifications Degree in Computer Science, Electrical Computer Engineering or similar field. 5+ years of experience in software development, with a focus on Functional Programming Scala stacks. Hands-on experience with AWS cloud services and Kubernetes for container orchestration. Strong understanding of RESTful APIs, Microservices Architecture, and Event-Driven Architecture. Strong problem-solving skills, attention to detail, and the ability to work in a collaborative, fast-paced environment. Experience with: - Functional Programming Languages - Scala (preferred) using ZIO, Cats, Cats Effect - Others considered - Modern RX Java, Kotlin, Clojure/Lisp, Haskell, OCaml, Elm, Typescript Effect AWS, Cloudformation, Terraform or similar technologies. Experience in Istio and Spinnaker Full software development lifecycle experience - including maintenance after release. Experience in Agile software development practices and version control systems (e.g., Git, Jira) Preferred Qualifications Solid understanding of distributed systems and how to design and scale them. Proficiency with CI/CD pipelines and modern development tools. Experience in building observability solutions using tools like Datadog, including tagging, metrics, and logging. The hiring range for this position in New York, NY is $148,700 to $199,400 per year. The base pay actually offered will take into account internal equity and also may vary depending on the candidate's geographic region, job-related knowledge, skills, and experience among other factors. A bonus and/or long-term incentive units may be provided as part of the compensation package, in addition to the full range of medical, financial, and/or other benefits, dependent on the level and position offered.
08/05/2026
Full time
Disney Entertainment and ESPN Product & Technology Technology is at the heart of Disney's past, present, and future. Disney Entertainment and ESPN Product & Technology is a global organization of engineers, product developers, designers, technologists, data scientists, and more - all working to build and advance the technological backbone for Disney's media business globally. The team marries technology with creativity to build world-class products, enhance storytelling, and drive velocity, innovation, and scalability for our businesses. We are Storytellers and Innovators. Creators and Builders. Entertainers and Engineers. We work with every part of The Walt Disney Company's media portfolio to advance the technological foundation and consumer media touch points serving millions of people around the world. Here are a few reasons why we think you'd love working here: Building the future of Disney's media: Our Technologists are designing and building the products and platforms that will power our media, advertising, and distribution businesses for years to come. Reach, Scale & Impact: More than ever, Disney's technology and products serve as a signature doorway for fans' connections with the company's brands and stories. Disney+. Hulu. ESPN. ABC. ABC News and many more. These products and brands - and the unmatched stories, storytellers, and events they carry - matter to millions of people globally. Innovation: We develop and implement groundbreaking products and techniques that shape industry norms, and solve complex and distinctive technical problems. Commerce, Data & Identity provides the core product management functions for areas crucial to Disney's media businesses. These include initiatives and products that power digital commerce, identity, and growth, as well as those that reach uniquely across The Walt Disney Company enterprise, such as messaging and privacy, among others. Additionally, it is responsible for the data engineering, science, and products for Disney Entertainment & ESPN, along with their interconnection with other parts of The Walt Disney Company. The Subscriptions team is looking to expand our team in the NYC office. We are a dynamic team of engineers that support all your Disney, ESPN and Hulu subscriptions at scale! We solve problems at internet-scale that only a handful of teams ever reach. We are seeking a highly skilled and motivated Senior Software Engineer to join our dynamic engineering team. In this role, you will be responsible for designing, developing, and deploying scalable and efficient software solutions using Functional Programming Style Scala and other backend languages. You will work with cutting-edge technologies in cloud computing, containerization, distributed systems, and observability. The ideal candidate should have a strong background in full-stack development, modern DevOps practices, and experience working with cloud-native technologies such as AWS and Kubernetes. Responsibilities Development: Design, develop, and maintain scalable, secure, and efficient software applications using Scala and other backend languages. Cloud Infrastructure: Utilize AWS services (CFN, EC2, Lambda, S3, Dynamo, etc.) to deploy and manage applications in the cloud. Containerization & Orchestration: Implement containerized solutions using Docker, deploy and manage services on Kubernetes. Distributed Systems: Design and build distributed systems that are fault-tolerant, highly available, and scalable. Understand concepts such as event-driven architecture, microservices, and data consistency. Observability & Monitoring: Implement and maintain observability best practices, including tagging, metrics, and logging to provide comprehensive visibility into system performance. Use tools like Datadog to monitor the health and performance of applications in real-time. Collaboration: Work closely with cross-functional teams including product managers, UI/UX designers, and DevOps engineers to define requirements, design systems, and deliver features. Code Quality & Best Practices: Write clean, maintainable, and testable code while following best practices for software development. Perform code reviews and provide mentorship to junior engineers on and offshore. Problem-Solving & Innovation: Tackle complex technical challenges and continuously seek opportunities to improve system performance, scalability, and reliability.AI Best Practices: Follow team's best practices regarding the safe use of LLMs and suggest improvements as appropriate. Basic Qualifications Degree in Computer Science, Electrical Computer Engineering or similar field. 5+ years of experience in software development, with a focus on Functional Programming Scala stacks. Hands-on experience with AWS cloud services and Kubernetes for container orchestration. Strong understanding of RESTful APIs, Microservices Architecture, and Event-Driven Architecture. Strong problem-solving skills, attention to detail, and the ability to work in a collaborative, fast-paced environment. Experience with: - Functional Programming Languages - Scala (preferred) using ZIO, Cats, Cats Effect - Others considered - Modern RX Java, Kotlin, Clojure/Lisp, Haskell, OCaml, Elm, Typescript Effect AWS, Cloudformation, Terraform or similar technologies. Experience in Istio and Spinnaker Full software development lifecycle experience - including maintenance after release. Experience in Agile software development practices and version control systems (e.g., Git, Jira) Preferred Qualifications Solid understanding of distributed systems and how to design and scale them. Proficiency with CI/CD pipelines and modern development tools. Experience in building observability solutions using tools like Datadog, including tagging, metrics, and logging. The hiring range for this position in New York, NY is $148,700 to $199,400 per year. The base pay actually offered will take into account internal equity and also may vary depending on the candidate's geographic region, job-related knowledge, skills, and experience among other factors. A bonus and/or long-term incentive units may be provided as part of the compensation package, in addition to the full range of medical, financial, and/or other benefits, dependent on the level and position offered.
Senior Cloud System Architect B-21 (Dayton)
DCS Corp Beavercreek, Ohio
Responsible for leading the creation of a technology framework and providing technical leadership to support cloud computing and automation efforts, with a focus on the design of systems and services that run on cloud platforms designed to support processing highly classified data. Essential Job Functions: Provide architectural leadership and best practices to cloud-based DevOps application development teams. Responsible for ensuring that critical applications are designed and optimized for high availability and disaster recovery. Strong leadership and team-building skills, and must be able to collaborate effectively with a group of high performing individuals. Support development and review of the full range of documentation and products required by DoDI 5000.02. Assist with execution of the USAF Airworthiness process per MIL-HDBK-516 and the AF Life-Cycle Systems Engineering (LCSE) and Operational Safety, Suitability, and Effectiveness (OSS&E) processes as prescribed by AFMCI 63-1201. Support implementation of Air Force Integrity Programs. Support requirements development, allocation, and verification. Support review and evaluation of weapon system design, analyses, and verification documentation and reports. Gather and develop technical and business data necessary to build and, as required, present program briefs. Support program meetings, working groups, and technical and business reviews as needed. Analyze equipment and software performance deficiencies and make recommendations for corrective actions. Support development of contract and documentation changes and updates. Support source selections in a technical advisory role as requested. Required Skills: Due to the sensitivity of customer related requirements, U.S. Citizenship is required. Must have bachelor's degree in Computer Science or a closely related subject. Have least 10 years' experience in designing large and complex IT operations in large organizations, research environments or academic environments. TS/SCI clearance is required. Desired Skills: A master's degree in Computer Science or a closely related subject is highly desired.
08/05/2026
Full time
Responsible for leading the creation of a technology framework and providing technical leadership to support cloud computing and automation efforts, with a focus on the design of systems and services that run on cloud platforms designed to support processing highly classified data. Essential Job Functions: Provide architectural leadership and best practices to cloud-based DevOps application development teams. Responsible for ensuring that critical applications are designed and optimized for high availability and disaster recovery. Strong leadership and team-building skills, and must be able to collaborate effectively with a group of high performing individuals. Support development and review of the full range of documentation and products required by DoDI 5000.02. Assist with execution of the USAF Airworthiness process per MIL-HDBK-516 and the AF Life-Cycle Systems Engineering (LCSE) and Operational Safety, Suitability, and Effectiveness (OSS&E) processes as prescribed by AFMCI 63-1201. Support implementation of Air Force Integrity Programs. Support requirements development, allocation, and verification. Support review and evaluation of weapon system design, analyses, and verification documentation and reports. Gather and develop technical and business data necessary to build and, as required, present program briefs. Support program meetings, working groups, and technical and business reviews as needed. Analyze equipment and software performance deficiencies and make recommendations for corrective actions. Support development of contract and documentation changes and updates. Support source selections in a technical advisory role as requested. Required Skills: Due to the sensitivity of customer related requirements, U.S. Citizenship is required. Must have bachelor's degree in Computer Science or a closely related subject. Have least 10 years' experience in designing large and complex IT operations in large organizations, research environments or academic environments. TS/SCI clearance is required. Desired Skills: A master's degree in Computer Science or a closely related subject is highly desired.
Systems Engineer, Journeyman - TS
DCS Corp Bedford, Massachusetts
s an exciting opportunity for a Systems Engineer to provide support to the Command, Control, Communications, and Battle Management Division (C3BM). Command, Control, Communications, and Battle Management (C3BM) has been tasked with delivering an integrated Department of the Air Force (DAF) Battle Network providing resilient decision advantage and enabling the USAF, USSF, Joint, and Coalition Force to win against the pacing challenge. C3BM supports execution in many different focus areas. C3BM's main efforts are Architecture and Systems Engineering (ASE), Operational Response Team (ORT), and multiple mission integration teams such as Air, Maritime, and multiple acquisitions consisting of both the Advanced Battle Management System (ABMS) and Space. The selected candidate will provide systems engineering support across the acquisition lifecycle and will be aligned to the Air Mission Integration Team (MIT). The MIT is responsible for translating operational requirements into technical solutions, developing integrated architectures, assessing and mitigating risks, supporting acquisition execution strategies, and coordinating test and evaluation activities to ensure capabilities are delivered to operational users. This is a full-time position located at Hanscom AFB, Bedford, MA and is 100% onsite. Essential Job Functions: This position provides the opportunity to work on some of the most dynamic, unique, and important programs supporting the United States Air Force. The selected candidate will work directly with senior leadership while supporting critical mission integration efforts that help shape the future of the DAF Battle Network and create opportunities for professional growth and development. Provide systems engineering support throughout the acquisition lifecycle, including requirements analysis, system design, integration, sustainment, and disposal activities. Support Mission Integration Team (MIT) activities by translating operational and functional requirements into technical requirements. Participate in architecture definition efforts to ensure integration and interoperability across the DAF C3BM enterprise architecture. Review current DoD architecture models and support development of future-state architectures that enable sensor-to-shooter connectivity. Support risk assessments and collaborate with stakeholders to identify and mitigate technical and programmatic risks. Assist in developing execution management strategies and transition capability requirements to the acquisition community. Support development of test and evaluation strategies in coordination with acquisition and test organizations. Drive interoperability and integration efforts across Program Executive Offices (PEOs) and weapon systems. Capture and analyze current and future operational architecture based on mission and engagement scenarios. Engage with joint and coalition stakeholders to support development of integrated multi-domain architectures. Support identification, assessment, and maturation of innovative concepts through analysis, modeling, simulation, and systems engineering activities. Prepare and support technical briefings, engineering documentation, reports, and recommendations for Government leadership. Required Skills: Due to the sensitivity of the customer, U.S. citizenship is required. Must have and be able to maintain an active Top Secret level clearance and be SCI eligible. Bachelor's or master's Degree in a related field, and 5 years of experience and within the DoD sector. Systems Engineering across the acquisition lifecycle. Systems Architecture and Integration. Requirements Analysis and Development. Mission Integration and System-of-Systems Engineering. Operational Analysis and Architecture Development. Cloud-based systems, including management and projection of cost and performance. Agile methodologies, CI/CD, DevSecOps, and DevOps principles. Knowledge of systems acquisition and program management processes as defined in DoDI 5000.02 and DoDI 5000.75. Modeling, Simulation, and Analysis. Technical Documentation and Brief Development. Travel may be required per the customer's discretion. Salary Range $96,757-$145,000 At DCS, we pride ourselves on providing flexibility that allows employees to balance meaningful work with their personal lives. We offer competitive compensation, benefits, and opportunities for learning and development. Our broad and competitive mix of benefits is designed to support and protect employees and their families. Our robust benefit offerings include medical, dental, 401k, ESOP, PTO, education reimbursement, work/life balance, parental and other leave programs. Learn more about our benefits here: DCS Corp Benefits
08/05/2026
Full time
s an exciting opportunity for a Systems Engineer to provide support to the Command, Control, Communications, and Battle Management Division (C3BM). Command, Control, Communications, and Battle Management (C3BM) has been tasked with delivering an integrated Department of the Air Force (DAF) Battle Network providing resilient decision advantage and enabling the USAF, USSF, Joint, and Coalition Force to win against the pacing challenge. C3BM supports execution in many different focus areas. C3BM's main efforts are Architecture and Systems Engineering (ASE), Operational Response Team (ORT), and multiple mission integration teams such as Air, Maritime, and multiple acquisitions consisting of both the Advanced Battle Management System (ABMS) and Space. The selected candidate will provide systems engineering support across the acquisition lifecycle and will be aligned to the Air Mission Integration Team (MIT). The MIT is responsible for translating operational requirements into technical solutions, developing integrated architectures, assessing and mitigating risks, supporting acquisition execution strategies, and coordinating test and evaluation activities to ensure capabilities are delivered to operational users. This is a full-time position located at Hanscom AFB, Bedford, MA and is 100% onsite. Essential Job Functions: This position provides the opportunity to work on some of the most dynamic, unique, and important programs supporting the United States Air Force. The selected candidate will work directly with senior leadership while supporting critical mission integration efforts that help shape the future of the DAF Battle Network and create opportunities for professional growth and development. Provide systems engineering support throughout the acquisition lifecycle, including requirements analysis, system design, integration, sustainment, and disposal activities. Support Mission Integration Team (MIT) activities by translating operational and functional requirements into technical requirements. Participate in architecture definition efforts to ensure integration and interoperability across the DAF C3BM enterprise architecture. Review current DoD architecture models and support development of future-state architectures that enable sensor-to-shooter connectivity. Support risk assessments and collaborate with stakeholders to identify and mitigate technical and programmatic risks. Assist in developing execution management strategies and transition capability requirements to the acquisition community. Support development of test and evaluation strategies in coordination with acquisition and test organizations. Drive interoperability and integration efforts across Program Executive Offices (PEOs) and weapon systems. Capture and analyze current and future operational architecture based on mission and engagement scenarios. Engage with joint and coalition stakeholders to support development of integrated multi-domain architectures. Support identification, assessment, and maturation of innovative concepts through analysis, modeling, simulation, and systems engineering activities. Prepare and support technical briefings, engineering documentation, reports, and recommendations for Government leadership. Required Skills: Due to the sensitivity of the customer, U.S. citizenship is required. Must have and be able to maintain an active Top Secret level clearance and be SCI eligible. Bachelor's or master's Degree in a related field, and 5 years of experience and within the DoD sector. Systems Engineering across the acquisition lifecycle. Systems Architecture and Integration. Requirements Analysis and Development. Mission Integration and System-of-Systems Engineering. Operational Analysis and Architecture Development. Cloud-based systems, including management and projection of cost and performance. Agile methodologies, CI/CD, DevSecOps, and DevOps principles. Knowledge of systems acquisition and program management processes as defined in DoDI 5000.02 and DoDI 5000.75. Modeling, Simulation, and Analysis. Technical Documentation and Brief Development. Travel may be required per the customer's discretion. Salary Range $96,757-$145,000 At DCS, we pride ourselves on providing flexibility that allows employees to balance meaningful work with their personal lives. We offer competitive compensation, benefits, and opportunities for learning and development. Our broad and competitive mix of benefits is designed to support and protect employees and their families. Our robust benefit offerings include medical, dental, 401k, ESOP, PTO, education reimbursement, work/life balance, parental and other leave programs. Learn more about our benefits here: DCS Corp Benefits
Senior DataStage ETL/Data Engineer
BC Forward Addison, Texas
Job Title: Senior DataStage ETL/Data Engineer Location: Addison, TX / Charlotte, NC / Newark, DE (Hybrid, minimum 3 days onsite) Duration: Contract - 17 months Approved locations: TX, Addison-$96.92 (Pay Rate: 68.25) NC, Charlotte-$92.39 (Pay Rate: 65.06) DE, Newark $92.39 (Pay Rate: 65.06) Job ID: 407307 About BCforward BCforward is a leading global IT consulting and workforce solutions firm providing services and support to Fortune 500 and government clients. Founded in 1998, BCforward has grown with our customers needs into a full-service business solutions provider. With delivery centers and offices across North America and India, we take pride in building long-term relationships and delivering excellence through innovation, collaboration, and integrity. Job Description We are seeking a Senior DataStage ETL/Data Engineer to join our dynamic team within I&DS Operational Datastores and Real-time Data Replication. The ideal candidate will have strong experience in IBM InfoSphere DataStage, IBM CloudPak for Data (CP4D) v4.5, Oracle, DB2, and advanced SQL/PLSQL and a proven ability to design, migrate, optimize, and support enterprise-scale data integration and transformation workloads with a focus on scalability, resiliency, performance, and data integrity. This role supports the 2026 Never Down Program and EOL technology migrations involving IBM CloudPak for Data. Responsibilities: Design, develop, and modernize data integration and transformation solutions using IBM DataStage and CP4D v4.5. Migrate and optimize critical data processing workloads across Oracle and DB2 platforms. Implement and support large-scale batch processing with robust performance, scalability, and resiliency. Perform advanced SQL/PLSQL development, data modeling, validation, and performance tuning. Collaborate with stakeholders to define requirements and ensure alignment with data governance and audit standards. Support real-time integration and CDC patterns for mission-critical data platforms. Required Skills & Qualifications: 12+ years in Data Engineering, ETL Development, or Data Integration. Strong expertise in IBM InfoSphere DataStage development. Hands-on experience with IBM CloudPak for Data (CP4D) v4.5. Advanced SQL and PL/SQL programming skills. Strong experience with Oracle and DB2 database platforms. Background in data transformation, data engineering, and data modeling. Experience designing and supporting large-scale batch processing environments. Troubleshooting, performance tuning, and data validation skills. Experience working in Linux/Unix environments. Excellent communication and stakeholder collaboration abilities. Familiarity with Change Data Capture (CDC) technologies. Preferred Skills: Experience in Agile development environments. Experience with cloud platforms and CI/CD pipelines with DevOps practices. Familiarity with data replication, CDC, and real-time integration technologies. Experience supporting mission-critical financial or enterprise data platforms. Knowledge of data governance, SOX controls, and audit requirements. Exposure to CockroachDB or MemSQL, Python-based scripting, Oracle GoldenGate, and GitHub Copilot. Additional Details: Start month: September. Interview process: 1-2 video rounds with Glider ID verification. Site codes and rates: TX8-044-02-18 Addison $96.93/hr; NC1-025-07-22 Charlotte $92.40/hr; DE5-021-03-03 Newark $92.40/hr. Maximum submissions per vendor: 3. Hours per day: 8. Labor type: Technical. Why BCforward? At BCforward, we believe in advancing lives and careers. When you join our team, you gain access to: Competitive compensation and benefits. Opportunities for growth with global clients. A supportive, inclusive culture that values innovation and people. Exposure to cutting-edge technologies and projects. About Our Commitment BCforward is an equal opportunity employer. We value diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, or veteran status. Interested? Apply Now! If this sounds like the right opportunity for you, please apply with your most recent resume. On your resume, include your current location or relocation plan, availability or start date, and whether a cooling-off period is required.
08/05/2026
Full time
Job Title: Senior DataStage ETL/Data Engineer Location: Addison, TX / Charlotte, NC / Newark, DE (Hybrid, minimum 3 days onsite) Duration: Contract - 17 months Approved locations: TX, Addison-$96.92 (Pay Rate: 68.25) NC, Charlotte-$92.39 (Pay Rate: 65.06) DE, Newark $92.39 (Pay Rate: 65.06) Job ID: 407307 About BCforward BCforward is a leading global IT consulting and workforce solutions firm providing services and support to Fortune 500 and government clients. Founded in 1998, BCforward has grown with our customers needs into a full-service business solutions provider. With delivery centers and offices across North America and India, we take pride in building long-term relationships and delivering excellence through innovation, collaboration, and integrity. Job Description We are seeking a Senior DataStage ETL/Data Engineer to join our dynamic team within I&DS Operational Datastores and Real-time Data Replication. The ideal candidate will have strong experience in IBM InfoSphere DataStage, IBM CloudPak for Data (CP4D) v4.5, Oracle, DB2, and advanced SQL/PLSQL and a proven ability to design, migrate, optimize, and support enterprise-scale data integration and transformation workloads with a focus on scalability, resiliency, performance, and data integrity. This role supports the 2026 Never Down Program and EOL technology migrations involving IBM CloudPak for Data. Responsibilities: Design, develop, and modernize data integration and transformation solutions using IBM DataStage and CP4D v4.5. Migrate and optimize critical data processing workloads across Oracle and DB2 platforms. Implement and support large-scale batch processing with robust performance, scalability, and resiliency. Perform advanced SQL/PLSQL development, data modeling, validation, and performance tuning. Collaborate with stakeholders to define requirements and ensure alignment with data governance and audit standards. Support real-time integration and CDC patterns for mission-critical data platforms. Required Skills & Qualifications: 12+ years in Data Engineering, ETL Development, or Data Integration. Strong expertise in IBM InfoSphere DataStage development. Hands-on experience with IBM CloudPak for Data (CP4D) v4.5. Advanced SQL and PL/SQL programming skills. Strong experience with Oracle and DB2 database platforms. Background in data transformation, data engineering, and data modeling. Experience designing and supporting large-scale batch processing environments. Troubleshooting, performance tuning, and data validation skills. Experience working in Linux/Unix environments. Excellent communication and stakeholder collaboration abilities. Familiarity with Change Data Capture (CDC) technologies. Preferred Skills: Experience in Agile development environments. Experience with cloud platforms and CI/CD pipelines with DevOps practices. Familiarity with data replication, CDC, and real-time integration technologies. Experience supporting mission-critical financial or enterprise data platforms. Knowledge of data governance, SOX controls, and audit requirements. Exposure to CockroachDB or MemSQL, Python-based scripting, Oracle GoldenGate, and GitHub Copilot. Additional Details: Start month: September. Interview process: 1-2 video rounds with Glider ID verification. Site codes and rates: TX8-044-02-18 Addison $96.93/hr; NC1-025-07-22 Charlotte $92.40/hr; DE5-021-03-03 Newark $92.40/hr. Maximum submissions per vendor: 3. Hours per day: 8. Labor type: Technical. Why BCforward? At BCforward, we believe in advancing lives and careers. When you join our team, you gain access to: Competitive compensation and benefits. Opportunities for growth with global clients. A supportive, inclusive culture that values innovation and people. Exposure to cutting-edge technologies and projects. About Our Commitment BCforward is an equal opportunity employer. We value diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, or veteran status. Interested? Apply Now! If this sounds like the right opportunity for you, please apply with your most recent resume. On your resume, include your current location or relocation plan, availability or start date, and whether a cooling-off period is required.
Senior Specialist AI & Data Solutions
BC Forward Lawrenceville, New Jersey
Job Title: Senior Specialist AI & Data Solutions Location: 50% onsite at Lawrence Township, NJ Duration: Temp - 12 months Pay Range: $38-48/hr (W2) Work Schedule: Mon-Fri, 8am-5pm About TSR TSR is a trusted staffing and workforce solutions partner with more than 50 years of experience delivery highly qualified talent to support clients' most critical business and technology initiatives. Through a disciplined approach to sourcing, candidate vetting, and delivery, TSR helps organizations scale teams quickly while maintaining quality and reliability. TSR operates with a global delivery model, leveraging specialized teams to support enterprise clients across North America and beyond. In June 2024, TSR was acquired by Justin Christian, founder and CEO of BCforward, a global provider of professional services and workforce solutions. This partnership expands TSR's ability to deliver broader capabilities, global delivery resources, and enhanced opportunities for both clients and consultants. Job Description We are seeking a Senior Specialist AI & Data Solutions to join our Global Digital Engagement Platforms team as the Web Delivery Lead. The ideal candidate will have strong experience in delivery leadership for Adobe Experience Manager (AEM), Python-based automation, AI/ML model integration, and prompt engineering and a proven ability to launch enterprise websites on time and within scope while embedding AI-powered capabilities across the web delivery lifecycle. Responsibilities: Lead end-to-end delivery of AEM-based websites across global markets, managing timelines, risks, and stakeholder expectations. Define and enforce delivery standards including intake workflows, QA processes, and go-live readiness criteria. Manage and mentor a distributed team of developers, content authors, and QA specialists. Ensure compliance with regulatory, accessibility (WCAG/EAA), cookie consent, and governance requirements. Maintain hands-on expertise across AEM Sites, Assets, and Cloud Service, guiding component development, template design, content fragment modeling, and multi-site management. Support adoption of headless CMS capabilities, GraphQL APIs, and omnichannel content delivery. Develop and maintain Python automation for site setup, content migration, metadata enrichment, QA validation, and reporting. Design, train, and fine-tune AI/ML models for automated tagging, content generation and summarization, predictive analytics, and image optimization. Lead prompt engineering initiatives for LLM-driven web copy, meta descriptions, SEO content, MLR package preparation, chatbots, and compliance pre-screening. Build and operationalize AI-powered workflows within AEM using APIs from OpenAI, Azure AI, AWS Bedrock, or similar platforms. Evaluate emerging AI tools and frameworks, prototype solutions, and present business cases for adoption with clear governance and benchmarks. Required Skills & Qualifications: Bachelor's or Master's degree in Computer Science, Engineering, or related field. 2-3+ years in web delivery or digital platforms. AEM expertise: Sites, Cloud Service, Java, OSGi, JCR, JavaScript, Content Fragments, and multi-site management. Python 3.x proficiency for automation scripting, API integrations, and ML model development using TensorFlow, PyTorch, or Hugging Face. AI/ML and prompt engineering experience with LLMs (e.g., GPT, Claude, LLaMA), prompt optimization, model fine-tuning, RAG patterns, and vector databases. Proven ability to lead globally distributed, cross-functional teams in agile environments with clear stakeholder communication. Working knowledge of Adobe Experience Cloud, cloud platforms (AWS, Azure, GCP), and DevOps tooling (CI/CD, Docker, Kubernetes). Familiarity with CDN, DNS, WAF, cookie consent, web accessibility (WCAG/EAA), and pharma or life sciences compliance requirements. Experience with NLP, personalization algorithms, semantic search, and evaluation of emerging AI technologies. Preferred Skills: Experience with headless CMS patterns in AEM and GraphQL APIs for omnichannel delivery. Background in building AI-enabled content operations and governance, including prompt libraries and model evaluation benchmarks. Why TSR? At TSR, we believe in advancing lives and careers. When you join our team, you gain access to: Competitive compensation and benefits. Opportunities for growth with global clients. A supportive, inclusive culture that values innovation and people. Exposure to cutting-edge technologies and projects. About Our Commitment TSR is an equal opportunity employer. We value diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, or veteran status. Interested? Apply Now! If this sounds like the right opportunity for you, please apply with your most recent resume
08/05/2026
Full time
Job Title: Senior Specialist AI & Data Solutions Location: 50% onsite at Lawrence Township, NJ Duration: Temp - 12 months Pay Range: $38-48/hr (W2) Work Schedule: Mon-Fri, 8am-5pm About TSR TSR is a trusted staffing and workforce solutions partner with more than 50 years of experience delivery highly qualified talent to support clients' most critical business and technology initiatives. Through a disciplined approach to sourcing, candidate vetting, and delivery, TSR helps organizations scale teams quickly while maintaining quality and reliability. TSR operates with a global delivery model, leveraging specialized teams to support enterprise clients across North America and beyond. In June 2024, TSR was acquired by Justin Christian, founder and CEO of BCforward, a global provider of professional services and workforce solutions. This partnership expands TSR's ability to deliver broader capabilities, global delivery resources, and enhanced opportunities for both clients and consultants. Job Description We are seeking a Senior Specialist AI & Data Solutions to join our Global Digital Engagement Platforms team as the Web Delivery Lead. The ideal candidate will have strong experience in delivery leadership for Adobe Experience Manager (AEM), Python-based automation, AI/ML model integration, and prompt engineering and a proven ability to launch enterprise websites on time and within scope while embedding AI-powered capabilities across the web delivery lifecycle. Responsibilities: Lead end-to-end delivery of AEM-based websites across global markets, managing timelines, risks, and stakeholder expectations. Define and enforce delivery standards including intake workflows, QA processes, and go-live readiness criteria. Manage and mentor a distributed team of developers, content authors, and QA specialists. Ensure compliance with regulatory, accessibility (WCAG/EAA), cookie consent, and governance requirements. Maintain hands-on expertise across AEM Sites, Assets, and Cloud Service, guiding component development, template design, content fragment modeling, and multi-site management. Support adoption of headless CMS capabilities, GraphQL APIs, and omnichannel content delivery. Develop and maintain Python automation for site setup, content migration, metadata enrichment, QA validation, and reporting. Design, train, and fine-tune AI/ML models for automated tagging, content generation and summarization, predictive analytics, and image optimization. Lead prompt engineering initiatives for LLM-driven web copy, meta descriptions, SEO content, MLR package preparation, chatbots, and compliance pre-screening. Build and operationalize AI-powered workflows within AEM using APIs from OpenAI, Azure AI, AWS Bedrock, or similar platforms. Evaluate emerging AI tools and frameworks, prototype solutions, and present business cases for adoption with clear governance and benchmarks. Required Skills & Qualifications: Bachelor's or Master's degree in Computer Science, Engineering, or related field. 2-3+ years in web delivery or digital platforms. AEM expertise: Sites, Cloud Service, Java, OSGi, JCR, JavaScript, Content Fragments, and multi-site management. Python 3.x proficiency for automation scripting, API integrations, and ML model development using TensorFlow, PyTorch, or Hugging Face. AI/ML and prompt engineering experience with LLMs (e.g., GPT, Claude, LLaMA), prompt optimization, model fine-tuning, RAG patterns, and vector databases. Proven ability to lead globally distributed, cross-functional teams in agile environments with clear stakeholder communication. Working knowledge of Adobe Experience Cloud, cloud platforms (AWS, Azure, GCP), and DevOps tooling (CI/CD, Docker, Kubernetes). Familiarity with CDN, DNS, WAF, cookie consent, web accessibility (WCAG/EAA), and pharma or life sciences compliance requirements. Experience with NLP, personalization algorithms, semantic search, and evaluation of emerging AI technologies. Preferred Skills: Experience with headless CMS patterns in AEM and GraphQL APIs for omnichannel delivery. Background in building AI-enabled content operations and governance, including prompt libraries and model evaluation benchmarks. Why TSR? At TSR, we believe in advancing lives and careers. When you join our team, you gain access to: Competitive compensation and benefits. Opportunities for growth with global clients. A supportive, inclusive culture that values innovation and people. Exposure to cutting-edge technologies and projects. About Our Commitment TSR is an equal opportunity employer. We value diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, or veteran status. Interested? Apply Now! If this sounds like the right opportunity for you, please apply with your most recent resume
Team Lead, Integrations
Genesis10 Irving, Texas
Genesis10 is currently seeking a Team Lead, Integrations for a direct hire opportunity with a Financial Services Corporation located in Irving, TX. The Lead Integration Engineer will oversee the day-to-day work of the integration engineering team and ensure solutions that are built match the respective designs. This role sits between the Integration Solutions Architect and the engineering team, where the Architect creates the design and the Lead is accountable for the quality, consistency, and completeness of the implementation. The ideal candidate is a strong .NET and Azure practitioner who still writes and reviews code regularly, but who is equally comfortable coaching engineers and holding the team to a high documentation standard. This is a working technical lead role, not a purely people-management role. Responsibilities: Oversee the work of Senior Integration Engineers and Integration Engineers, assigning and sequencing implementation work against architectural designs and sprint commitments Track team throughput and quality - surfacing blockers, skill gaps, or capacity risks to the Architect and IT leadership Mentor engineers on .NET craftsmanship, Azure integration patterns, and troubleshooting technique; run knowledge-sharing sessions on recurring issues Act as the first escalation point for implementation-level technical questions before they go to the Architect Perform code reviews on all integration/engineering work prior to merge/release, checking for correctness, security, performance, and adherence to team coding standards Serve as the quality gate between engineering output and production Identify recurring code-quality issues across the team and turn them into coding standards, checklists, or automated lint/analyzer rules Verify that implementations correctly apply the integration design patterns specified by the architecture Maintain a living reference of approved pattern implementations (reference code, snippets, examples) the team can build from consistently Ensure every integration solution ships with complete, audit-ready documentation Define and enforce the team's documentation checklist/definition-of-done Keep architecture and Lucid chart diagrams synchronized with what is deployed Continue to design and write .NET code and Azure configuration directly for complex, high-risk, or exemplar pieces of work Lead root-cause investigation on significant production incidents Partner with the Architect on estimation, technical risk assessment, and sequencing for upcoming integration initiatives Contribute to and enforce infrastructure-as-code and CI/CD standards Requirements: Leadership & Quality Assurance: 2 years of Team Leadership experience formally or informally leading engineers, assigning work, reviewing output, and being accountable for team quality Demonstrated track record running code reviews at scale, including giving direct, actionable feedback to engineers Experience defining or enforcing a definition-of-done that includes documentation, not just working code Integration & Messaging: Deep, hands-on experience with Azure Service Bus, Event Grid, and/or Event Hubs Configuring, securing, and reviewing APIs built by others within Azure API Management (APIM) Building and reviewing orchestration workflows in Azure Logic Apps Expert-level fluency with Integration Design Patterns Core Azure Platform: Container Apps, App Service, Azure Functions, Virtual Machines Virtual Networks (VNets), Private Endpoints, NSGs, Application Gateway, Azure Front Door Azure Blob Storage, Azure Files, Cosmos DB Data & Databases: Azure SQL, including reviewing query and schema design decisions Cosmos NoSQL DB Azure Cache for Redis Security & Identity: System- and user-assigned Managed Identity Secrets & Key Management: Azure Key Vault integration review Network Security: Private Link, Private Endpoints, WAF Identity & Access: App Registrations Authentication & Authorization: OAuth flows Infrastructure as Code & DevOps: IaC with Terraform; able to review and approve infrastructure changes Azure DevOps pipeline design, branch policies, and mandatory-review gate configuration Observability & Monitoring: Application Monitoring using Application Insights Log Management and Analytics Platform Monitoring using Azure Monitor Query & Analysis using KQL (Kusto Query Language) Alerting, dashboards, and validating solution observability Application Development: Minimum 8 years of solid .NET coding experience (C#, ASP.NET Core, background/worker services) Expert knowledge of common design and patterns including .NET patterns, libraries and Azure services Documentation & Communication: Strong technical writing skills Visual Modeling with Lucid chart / Azure architecture diagramming Comfortable delivering direct feedback in code review and 1:1 settings Experience working with Jira, including grooming and sequencing work for a team Demonstrated Track Record & Mindset: Accountability: Treats the team's output as their own Consistency: Applies review standards evenly across engineers Coaching Orientation: Uses code review as a teaching moment Escalation Judgment: Knows the difference between an implementation fix and a design flaw Experience & Education: Bachelor's degree in Computer Science, Information Systems, or equivalent practical experience 8 years of professional software engineering experience, including 3 years focused on cloud-based integration work Prior experience reviewing or leading the work of other engineers required Desired skills: Graph Database: Experience with graph database design and implementation (e.g., Azure Cosmos DB for Apache Gremlin) Data Integration: Azure Data Factory Analytics Platform: Microsoft Fabric Microsoft Power Platform: Power Apps, Power Automate, Power Pages, Dataverse Regulated Industry Experience: Prior experience in financial services or another regulated industry (GLBA, SOX, PCI-DSS, SOC 2) If you have the described qualifications and are interested in this exciting opportunity, please apply! Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals. For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10: Access to hundreds of clients, most who have been working with Genesis10 for 5-20 years. The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years. Access to an experienced, caring recruiting team (more than 7 years of experience, on average.) Behavioral Health Platform Medical, Dental, Vision Health Savings Account Voluntary Hospital Indemnity (Critical Illness & Accident) Voluntary Term Life Insurance 401K Sick Pay (for applicable states/municipalities) Commuter Benefits (Dallas, NYC, SF, and Illinois) For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website. Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
08/05/2026
Full time
Genesis10 is currently seeking a Team Lead, Integrations for a direct hire opportunity with a Financial Services Corporation located in Irving, TX. The Lead Integration Engineer will oversee the day-to-day work of the integration engineering team and ensure solutions that are built match the respective designs. This role sits between the Integration Solutions Architect and the engineering team, where the Architect creates the design and the Lead is accountable for the quality, consistency, and completeness of the implementation. The ideal candidate is a strong .NET and Azure practitioner who still writes and reviews code regularly, but who is equally comfortable coaching engineers and holding the team to a high documentation standard. This is a working technical lead role, not a purely people-management role. Responsibilities: Oversee the work of Senior Integration Engineers and Integration Engineers, assigning and sequencing implementation work against architectural designs and sprint commitments Track team throughput and quality - surfacing blockers, skill gaps, or capacity risks to the Architect and IT leadership Mentor engineers on .NET craftsmanship, Azure integration patterns, and troubleshooting technique; run knowledge-sharing sessions on recurring issues Act as the first escalation point for implementation-level technical questions before they go to the Architect Perform code reviews on all integration/engineering work prior to merge/release, checking for correctness, security, performance, and adherence to team coding standards Serve as the quality gate between engineering output and production Identify recurring code-quality issues across the team and turn them into coding standards, checklists, or automated lint/analyzer rules Verify that implementations correctly apply the integration design patterns specified by the architecture Maintain a living reference of approved pattern implementations (reference code, snippets, examples) the team can build from consistently Ensure every integration solution ships with complete, audit-ready documentation Define and enforce the team's documentation checklist/definition-of-done Keep architecture and Lucid chart diagrams synchronized with what is deployed Continue to design and write .NET code and Azure configuration directly for complex, high-risk, or exemplar pieces of work Lead root-cause investigation on significant production incidents Partner with the Architect on estimation, technical risk assessment, and sequencing for upcoming integration initiatives Contribute to and enforce infrastructure-as-code and CI/CD standards Requirements: Leadership & Quality Assurance: 2 years of Team Leadership experience formally or informally leading engineers, assigning work, reviewing output, and being accountable for team quality Demonstrated track record running code reviews at scale, including giving direct, actionable feedback to engineers Experience defining or enforcing a definition-of-done that includes documentation, not just working code Integration & Messaging: Deep, hands-on experience with Azure Service Bus, Event Grid, and/or Event Hubs Configuring, securing, and reviewing APIs built by others within Azure API Management (APIM) Building and reviewing orchestration workflows in Azure Logic Apps Expert-level fluency with Integration Design Patterns Core Azure Platform: Container Apps, App Service, Azure Functions, Virtual Machines Virtual Networks (VNets), Private Endpoints, NSGs, Application Gateway, Azure Front Door Azure Blob Storage, Azure Files, Cosmos DB Data & Databases: Azure SQL, including reviewing query and schema design decisions Cosmos NoSQL DB Azure Cache for Redis Security & Identity: System- and user-assigned Managed Identity Secrets & Key Management: Azure Key Vault integration review Network Security: Private Link, Private Endpoints, WAF Identity & Access: App Registrations Authentication & Authorization: OAuth flows Infrastructure as Code & DevOps: IaC with Terraform; able to review and approve infrastructure changes Azure DevOps pipeline design, branch policies, and mandatory-review gate configuration Observability & Monitoring: Application Monitoring using Application Insights Log Management and Analytics Platform Monitoring using Azure Monitor Query & Analysis using KQL (Kusto Query Language) Alerting, dashboards, and validating solution observability Application Development: Minimum 8 years of solid .NET coding experience (C#, ASP.NET Core, background/worker services) Expert knowledge of common design and patterns including .NET patterns, libraries and Azure services Documentation & Communication: Strong technical writing skills Visual Modeling with Lucid chart / Azure architecture diagramming Comfortable delivering direct feedback in code review and 1:1 settings Experience working with Jira, including grooming and sequencing work for a team Demonstrated Track Record & Mindset: Accountability: Treats the team's output as their own Consistency: Applies review standards evenly across engineers Coaching Orientation: Uses code review as a teaching moment Escalation Judgment: Knows the difference between an implementation fix and a design flaw Experience & Education: Bachelor's degree in Computer Science, Information Systems, or equivalent practical experience 8 years of professional software engineering experience, including 3 years focused on cloud-based integration work Prior experience reviewing or leading the work of other engineers required Desired skills: Graph Database: Experience with graph database design and implementation (e.g., Azure Cosmos DB for Apache Gremlin) Data Integration: Azure Data Factory Analytics Platform: Microsoft Fabric Microsoft Power Platform: Power Apps, Power Automate, Power Pages, Dataverse Regulated Industry Experience: Prior experience in financial services or another regulated industry (GLBA, SOX, PCI-DSS, SOC 2) If you have the described qualifications and are interested in this exciting opportunity, please apply! Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals. For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10: Access to hundreds of clients, most who have been working with Genesis10 for 5-20 years. The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years. Access to an experienced, caring recruiting team (more than 7 years of experience, on average.) Behavioral Health Platform Medical, Dental, Vision Health Savings Account Voluntary Hospital Indemnity (Critical Illness & Accident) Voluntary Term Life Insurance 401K Sick Pay (for applicable states/municipalities) Commuter Benefits (Dallas, NYC, SF, and Illinois) For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website. Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
Network Voice Engineer
Lennar Homes Miami, Florida
Network & Telecom Support Engineer II Summary of Position: The Network & Telecom Support Engineer II is responsible for the support, implementation, and maintenance of the organization's network infrastructure and telephony systems. This entry-level role assists in the deployment and optimization of networking technologies, including routers, switches, firewalls, and wireless access points, while also managing and supporting cloud-based telephony platforms such as Cisco Webex Cloud, Webex Contact Center, and Vonage Contact Center. The engineer will work collaboratively with senior engineers and cross-functional teams to ensure the reliable, secure, and efficient operation of both network and telecom environments that support the business. Principal Duties and Responsibilities: Network Implementation: Assist in the deployment of network infrastructure, including routers, switches, firewalls, and wireless access points. Participate in the implementation and configuration of network solutions, ensuring alignment with organizational standards and best practices. Deploy network configuration through automated workflows following CI/CD practices. Contribute to the deployment of Software-Defined Wide Area Network (SD-WAN) solutions to improve network efficiency and reliability. Assist in the implementation of cloud connectivity solutions, optimizing for performance, security, and cost-effectiveness. Telephony Management: Assist in operating and optimizing cloud-based telephony systems, including Webex Cloud for internal telephony, Webex Contact Center for IT Service Desk operations, and Vonage Contact Center for sales and customer care. Support the configuration and administration of telephony systems, leveraging automation tools and DevOps principles where applicable. Assist in the integration of telephony solutions with other IT systems and platforms. Monitor telephony system performance and usage metrics, analyzing data to ensure optimal performance and identify areas for improvement. Assist with telephony-related projects, including system upgrades, migrations, and new implementations. Operations and Support: Provide on-call service 24x7 for Tier 2 support in partnership with Service Desk for network and telephony-related issues, troubleshooting and resolving problems in a timely manner. Monitor network and telephony performance using monitoring tools to identify and address potential issues proactively. Assist in the administration of network and telephony security services, including updating rules and filters to maintain a strong security posture. Support infrastructure maintenance activities, including upgrades and vulnerability patches. Collaborate with vendors and PSTN partners to resolve issues, manage upgrades, and maintain service contracts. Process Improvement and Documentation: Contribute to the development and enforcement of Standard Operating Procedures (SOPs) for network and telephony operations. Identify opportunities for operational process improvements, supporting initiatives to enhance efficiency and service quality. Assist in maintaining up-to-date documentation, including network diagrams, telephony configurations, and standard operating procedures. Develop and deliver training materials and documentation for end-users on telephony and network systems. Education and Experience Requirements: Education: Bachelor's degree in Computer Science, Information Technology, Telecommunications, Engineering, or a related field; or equivalent work experience. Experience: 1-2 years of relevant work experience in computer networks, telephony, or related IT support roles. Exposure to enterprise office and data center environments is a plus. Certifications: Certifications such as ITIL, Cisco CCNA, CompTIA Network+, or similar are preferred but not required. Skills and Expertise: Workflow Automation: Identifying repetitive, rule-based tasks and leveraging AI to execute smart automations that save time and manual workload. Networking Technologies: Basic proficiency in configuring and managing network devices, including routers, switches, and firewalls. Familiarity with Cisco/Meraki, Cisco Catalyst, Cisco Nexus, and Palo Alto Networks is a plus. Telephony Systems: Foundational understanding of cloud-based telephony systems, including Cisco Webex Cloud and VoIP/SIP protocols. Experience with contact center platforms is a plus. Security: Knowledge of network and telephony security best practices, including compliance with security policies and industry regulations. Technical Proficiency: Familiarity with automation tools, GitHub, and CI/CD practices is beneficial. Troubleshooting: Strong problem-solving and analytical skills, with the ability to diagnose and resolve network and telephony issues efficiently. Personal Attributes: Team Player: Ability to work collaboratively with senior engineers, IT teams, vendors, and other stakeholders to achieve shared goals. Communication: Effective written and verbal communication skills, with the ability to explain technical concepts to non-technical audiences. Customer Focus: A commitment to providing high-quality support and improving the user experience for both internal and external stakeholders. Detail-Oriented: Strong attention to detail, ensuring accuracy in configurations and documentation. Adaptability: Ability to adapt to new technologies and processes in a dynamic work environment. Additional Requirements: Continuous Learning: Commitment to staying current with industry trends and pursuing relevant certifications and training opportunities. Travel: Willingness to travel occasionally to support network or telephony installations and upgrades at remote locations. This role is ideal for a motivated individual looking to build a foundation in both network engineering and telecommunications, gaining hands-on experience across infrastructure, cloud connectivity, and enterprise telephony systems. Physical Requirements This is primarily a sedentary office position which requires the incumbent to have the ability to operate computer equipment, speak, hear, bend, stoop, reach, lift, and move and carry up to 25 lbs. Finger dexterity is necessary. This description outlines the basic responsibilities and requirements for the position noted. This is not a comprehensive listing of all job duties of the associates. Duties, responsibilities, and activities may change at any time with or without notice. Life at Lennar At Lennar, we are committed to fostering a supportive and enriching environment for our Associates, offering a comprehensive array of benefits designed to enhance their well-being and professional growth. Our Associates have access to robust health insurance plans, including Medical, Dental, and Vision coverage, ensuring their health needs are well taken care of. Our 401(k) Retirement Plan, complete with a $1 for $1 Company Match up to 5%, helps secure their financial future, while Paid Parental Leave and an Associate Assistance Plan provide essential support during life's critical moments. To further support our Associates, we provide an Education Assistance Program and up to $30,000 in Adoption Assistance, underscoring our commitment to their diverse needs and aspirations. From the moment of hire, they can enjoy up to three weeks of vacation annually, alongside generous Holiday, Sick Leave, and Personal Day policies. Additionally, we offer a New Hire Referral Bonus Program, significant Home Purchase Discounts, and unique opportunities such as the Everyone's Included Day. At Lennar, we believe in investing in our Associates, empowering them to thrive both personally and professionally. Lennar Associates will have access to these benefits as outlined by Lennar's policies and applicable plan terms. Visit to view our suite of benefits. Join the fun and follow us on social media to see what's happening at our company, and don't forget to connect with us on Lennar: Overview LinkedIn for the latest job opportunities. Lennar is an equal opportunity employer and complies with all applicable federal, state, and local fair employment practices laws.
08/05/2026
Full time
Network & Telecom Support Engineer II Summary of Position: The Network & Telecom Support Engineer II is responsible for the support, implementation, and maintenance of the organization's network infrastructure and telephony systems. This entry-level role assists in the deployment and optimization of networking technologies, including routers, switches, firewalls, and wireless access points, while also managing and supporting cloud-based telephony platforms such as Cisco Webex Cloud, Webex Contact Center, and Vonage Contact Center. The engineer will work collaboratively with senior engineers and cross-functional teams to ensure the reliable, secure, and efficient operation of both network and telecom environments that support the business. Principal Duties and Responsibilities: Network Implementation: Assist in the deployment of network infrastructure, including routers, switches, firewalls, and wireless access points. Participate in the implementation and configuration of network solutions, ensuring alignment with organizational standards and best practices. Deploy network configuration through automated workflows following CI/CD practices. Contribute to the deployment of Software-Defined Wide Area Network (SD-WAN) solutions to improve network efficiency and reliability. Assist in the implementation of cloud connectivity solutions, optimizing for performance, security, and cost-effectiveness. Telephony Management: Assist in operating and optimizing cloud-based telephony systems, including Webex Cloud for internal telephony, Webex Contact Center for IT Service Desk operations, and Vonage Contact Center for sales and customer care. Support the configuration and administration of telephony systems, leveraging automation tools and DevOps principles where applicable. Assist in the integration of telephony solutions with other IT systems and platforms. Monitor telephony system performance and usage metrics, analyzing data to ensure optimal performance and identify areas for improvement. Assist with telephony-related projects, including system upgrades, migrations, and new implementations. Operations and Support: Provide on-call service 24x7 for Tier 2 support in partnership with Service Desk for network and telephony-related issues, troubleshooting and resolving problems in a timely manner. Monitor network and telephony performance using monitoring tools to identify and address potential issues proactively. Assist in the administration of network and telephony security services, including updating rules and filters to maintain a strong security posture. Support infrastructure maintenance activities, including upgrades and vulnerability patches. Collaborate with vendors and PSTN partners to resolve issues, manage upgrades, and maintain service contracts. Process Improvement and Documentation: Contribute to the development and enforcement of Standard Operating Procedures (SOPs) for network and telephony operations. Identify opportunities for operational process improvements, supporting initiatives to enhance efficiency and service quality. Assist in maintaining up-to-date documentation, including network diagrams, telephony configurations, and standard operating procedures. Develop and deliver training materials and documentation for end-users on telephony and network systems. Education and Experience Requirements: Education: Bachelor's degree in Computer Science, Information Technology, Telecommunications, Engineering, or a related field; or equivalent work experience. Experience: 1-2 years of relevant work experience in computer networks, telephony, or related IT support roles. Exposure to enterprise office and data center environments is a plus. Certifications: Certifications such as ITIL, Cisco CCNA, CompTIA Network+, or similar are preferred but not required. Skills and Expertise: Workflow Automation: Identifying repetitive, rule-based tasks and leveraging AI to execute smart automations that save time and manual workload. Networking Technologies: Basic proficiency in configuring and managing network devices, including routers, switches, and firewalls. Familiarity with Cisco/Meraki, Cisco Catalyst, Cisco Nexus, and Palo Alto Networks is a plus. Telephony Systems: Foundational understanding of cloud-based telephony systems, including Cisco Webex Cloud and VoIP/SIP protocols. Experience with contact center platforms is a plus. Security: Knowledge of network and telephony security best practices, including compliance with security policies and industry regulations. Technical Proficiency: Familiarity with automation tools, GitHub, and CI/CD practices is beneficial. Troubleshooting: Strong problem-solving and analytical skills, with the ability to diagnose and resolve network and telephony issues efficiently. Personal Attributes: Team Player: Ability to work collaboratively with senior engineers, IT teams, vendors, and other stakeholders to achieve shared goals. Communication: Effective written and verbal communication skills, with the ability to explain technical concepts to non-technical audiences. Customer Focus: A commitment to providing high-quality support and improving the user experience for both internal and external stakeholders. Detail-Oriented: Strong attention to detail, ensuring accuracy in configurations and documentation. Adaptability: Ability to adapt to new technologies and processes in a dynamic work environment. Additional Requirements: Continuous Learning: Commitment to staying current with industry trends and pursuing relevant certifications and training opportunities. Travel: Willingness to travel occasionally to support network or telephony installations and upgrades at remote locations. This role is ideal for a motivated individual looking to build a foundation in both network engineering and telecommunications, gaining hands-on experience across infrastructure, cloud connectivity, and enterprise telephony systems. Physical Requirements This is primarily a sedentary office position which requires the incumbent to have the ability to operate computer equipment, speak, hear, bend, stoop, reach, lift, and move and carry up to 25 lbs. Finger dexterity is necessary. This description outlines the basic responsibilities and requirements for the position noted. This is not a comprehensive listing of all job duties of the associates. Duties, responsibilities, and activities may change at any time with or without notice. Life at Lennar At Lennar, we are committed to fostering a supportive and enriching environment for our Associates, offering a comprehensive array of benefits designed to enhance their well-being and professional growth. Our Associates have access to robust health insurance plans, including Medical, Dental, and Vision coverage, ensuring their health needs are well taken care of. Our 401(k) Retirement Plan, complete with a $1 for $1 Company Match up to 5%, helps secure their financial future, while Paid Parental Leave and an Associate Assistance Plan provide essential support during life's critical moments. To further support our Associates, we provide an Education Assistance Program and up to $30,000 in Adoption Assistance, underscoring our commitment to their diverse needs and aspirations. From the moment of hire, they can enjoy up to three weeks of vacation annually, alongside generous Holiday, Sick Leave, and Personal Day policies. Additionally, we offer a New Hire Referral Bonus Program, significant Home Purchase Discounts, and unique opportunities such as the Everyone's Included Day. At Lennar, we believe in investing in our Associates, empowering them to thrive both personally and professionally. Lennar Associates will have access to these benefits as outlined by Lennar's policies and applicable plan terms. Visit to view our suite of benefits. Join the fun and follow us on social media to see what's happening at our company, and don't forget to connect with us on Lennar: Overview LinkedIn for the latest job opportunities. Lennar is an equal opportunity employer and complies with all applicable federal, state, and local fair employment practices laws.
Senior Cloud Platform Engineer (Kubernetes, AWS/GCP, Terraform)
HTC Global Services Inc Dearborn, Michigan
Job Title Senior Cloud Platform Engineer (Kubernetes, AWS/GCP, Terraform) Overview We are seeking a Senior Cloud Platform Engineer to design, develop, and maintain cloud platform infrastructure and automation solutions. This role focuses on building scalable, resilient, cloud-native platforms while improving infrastructure automation, deployment reliability, and operational excellence across multi-cloud environments. This is a hybrid position requiring four days per week in the office. Key Responsibilities Design and build cloud-native platform solutions and automation tooling. Develop automated infrastructure provisioning workflows. Build and maintain Infrastructure as Code using Terraform. Manage GitOps-driven infrastructure deployment workflows. Provision, manage, and automate Kubernetes clusters. Create and maintain Helm charts for Kubernetes deployments. Configure and manage service mesh technologies, including traffic management and security. Build and operate cloud platform services across AWS and GCP environments. Configure and maintain API gateway and ingress infrastructure. Develop and maintain CI/CD pipelines. Automate deployment validation, rollback, and migration processes. Implement cloud security best practices, including IAM and workload identity. Build monitoring, alerting, and observability solutions. Develop automation tools and scripts using Python, Go, Bash, or Node.js. Participate in on-call support and incident response. Troubleshoot production infrastructure issues and drive long-term reliability improvements. Required Qualifications Bachelor's or Master's degree in Computer Science, Engineering, or a related field. 5+ years of professional experience in Cloud Infrastructure, Platform Engineering, DevOps, or Site Reliability Engineering. Experience provisioning and managing production Kubernetes clusters (EKS or GKE). 3+ years of experience managing cloud infrastructure on AWS and/or GCP. Strong experience developing Infrastructure as Code using Terraform. Experience with GitOps workflows. Experience creating and maintaining Helm charts. Strong troubleshooting and problem-solving skills in cloud infrastructure and distributed systems. Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, or similar. Experience developing CI/CD pipelines. Strong infrastructure automation and scripting skills. Preferred Qualifications Experience with GKE, Pub/Sub, Cloud Storage, Workload Identity, VPC-SC, and Cloud Operations. Experience with API gateway technologies such as Tyk, Apigee, or NGINX. Experience with Istio service mesh. Experience with ArgoCD, Tekton, Concourse, or similar GitOps platforms. Programming experience with Python, Go, Java (Spring Boot), or Node.js. Experience with Atlantis, Terragrunt, or similar Terraform collaboration tools. Experience implementing cloud security best practices, including IAM, mTLS, OAuth2/OIDC, and multi-cloud security. Experience with gRPC services and Protocol Buffers. What Makes HTC A Great Place To Build Your Future HTC Global Services wants you to join our team. Come build new things with us and advance your career. At HTC Global, you'll collaborate with experts, work alongside clients, and be part of high-performing teams driving success together. You'll have long-term opportunities to grow your career and develop skills in the latest emerging technologies. At HTC Global Services, our employees have access to a comprehensive benefits package. Benefits can include Group Health (Medical, Dental, and Vision), Paid Time Off, Paid Holidays, 401(k) matching, Group Life and Disability insurance, Professional Development opportunities, Wellness programs, and a variety of other perks. Our success as a company is built on inclusion and diversity. HTC Global Services is committed to providing a workplace free from discrimination and harassment, where every employee is treated with dignity and respect. We celebrate differences and believe that diverse cultures, perspectives, and skills drive innovation and success. HTC is an Equal Opportunity Employer and a proud National Minority Supplier. We seek to empower each individual, fostering an environment where everyone feels valued, included, and respected.
08/05/2026
Full time
Job Title Senior Cloud Platform Engineer (Kubernetes, AWS/GCP, Terraform) Overview We are seeking a Senior Cloud Platform Engineer to design, develop, and maintain cloud platform infrastructure and automation solutions. This role focuses on building scalable, resilient, cloud-native platforms while improving infrastructure automation, deployment reliability, and operational excellence across multi-cloud environments. This is a hybrid position requiring four days per week in the office. Key Responsibilities Design and build cloud-native platform solutions and automation tooling. Develop automated infrastructure provisioning workflows. Build and maintain Infrastructure as Code using Terraform. Manage GitOps-driven infrastructure deployment workflows. Provision, manage, and automate Kubernetes clusters. Create and maintain Helm charts for Kubernetes deployments. Configure and manage service mesh technologies, including traffic management and security. Build and operate cloud platform services across AWS and GCP environments. Configure and maintain API gateway and ingress infrastructure. Develop and maintain CI/CD pipelines. Automate deployment validation, rollback, and migration processes. Implement cloud security best practices, including IAM and workload identity. Build monitoring, alerting, and observability solutions. Develop automation tools and scripts using Python, Go, Bash, or Node.js. Participate in on-call support and incident response. Troubleshoot production infrastructure issues and drive long-term reliability improvements. Required Qualifications Bachelor's or Master's degree in Computer Science, Engineering, or a related field. 5+ years of professional experience in Cloud Infrastructure, Platform Engineering, DevOps, or Site Reliability Engineering. Experience provisioning and managing production Kubernetes clusters (EKS or GKE). 3+ years of experience managing cloud infrastructure on AWS and/or GCP. Strong experience developing Infrastructure as Code using Terraform. Experience with GitOps workflows. Experience creating and maintaining Helm charts. Strong troubleshooting and problem-solving skills in cloud infrastructure and distributed systems. Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, or similar. Experience developing CI/CD pipelines. Strong infrastructure automation and scripting skills. Preferred Qualifications Experience with GKE, Pub/Sub, Cloud Storage, Workload Identity, VPC-SC, and Cloud Operations. Experience with API gateway technologies such as Tyk, Apigee, or NGINX. Experience with Istio service mesh. Experience with ArgoCD, Tekton, Concourse, or similar GitOps platforms. Programming experience with Python, Go, Java (Spring Boot), or Node.js. Experience with Atlantis, Terragrunt, or similar Terraform collaboration tools. Experience implementing cloud security best practices, including IAM, mTLS, OAuth2/OIDC, and multi-cloud security. Experience with gRPC services and Protocol Buffers. What Makes HTC A Great Place To Build Your Future HTC Global Services wants you to join our team. Come build new things with us and advance your career. At HTC Global, you'll collaborate with experts, work alongside clients, and be part of high-performing teams driving success together. You'll have long-term opportunities to grow your career and develop skills in the latest emerging technologies. At HTC Global Services, our employees have access to a comprehensive benefits package. Benefits can include Group Health (Medical, Dental, and Vision), Paid Time Off, Paid Holidays, 401(k) matching, Group Life and Disability insurance, Professional Development opportunities, Wellness programs, and a variety of other perks. Our success as a company is built on inclusion and diversity. HTC Global Services is committed to providing a workplace free from discrimination and harassment, where every employee is treated with dignity and respect. We celebrate differences and believe that diverse cultures, perspectives, and skills drive innovation and success. HTC is an Equal Opportunity Employer and a proud National Minority Supplier. We seek to empower each individual, fostering an environment where everyone feels valued, included, and respected.

Modal Window

  • Home
  • Contact
  • About Us
  • FAQs
  • Terms & Conditions
  • Privacy
  • Employer
  • Post a Job
  • Search Resumes
  • Sign in
  • Job Seeker
  • Find Jobs
  • Create Resume
  • Sign in
  • IT blog
  • Facebook
  • Twitter
  • LinkedIn
  • Youtube
© 2008-2026 IT Job Board