Engineering Operation Technician

  • Amazon Web Services, Inc.
  • Berwick, Pennsylvania
  • 09/11/2026
Full time Information Technology Telecommunications

Job Description

Amazon Web Services, Inc. seeks an Engineering Operation Technician to support and improve AWS infrastructure in a fast-paced, customer-obsessed environment. You will monitor and maintain data center and cloud systems, troubleshoot hardware and network issues, and respond to incidents to ensure high availability and performance. Responsibilities include executing changes, automating routine tasks with scripts, maintaining documentation and runbooks, and collaborating with engineering teams on root-cause analysis and continuous improvement. This role offers hands-on exposure to cutting-edge AWS services, large-scale distributed systems, and strong opportunities for growth and learning.

Responsibilities

  • Monitor and maintain AWS data center and cloud infrastructure for availability and performance.
  • Troubleshoot and resolve hardware, OS, and network incidents within defined SLAs.
  • Execute planned changes, maintenance, and deployments following change management processes.
  • Automate routine operational tasks using scripts and tooling to improve efficiency and reliability.
  • Use monitoring and alerting tools to detect, triage, and remediate system issues.
  • Participate in on-call rotations, incident response, and post-incident reviews with engineering teams.
  • Document procedures, runbooks, and system configurations for repeatable operations.
  • Collaborate with cross-functional AWS teams to support capacity planning and infrastructure upgrades.
  • Follow security, compliance, and safety standards across data center and cloud operations.
  • Contribute to continuous improvement initiatives to enhance system resilience and operational excellence.

Required Skills

  • Linux system administration
  • Network troubleshooting (TCP/IP, DNS, VPN)
  • Scripting (Python, Bash, Power
  • Shell)
  • Monitoring and alerting tools (Cloud
  • Watch, Prometheus, Grafana)
  • Incident and change management (ITIL)
  • Hardware and data center operations
  • Cloud infrastructure (AWS services)
  • Configuration management tools (Ansible, Chef, Puppet)
  • Virtualization and containers (EC2, Docker, Kubernetes)
  • Documentation and runbook creation