Back to Jobs
EverOps
Development 5h ago

Lead DevOps Engineer (Remote)

EverOps
United StatesUnited States
Full-time
Not Disclosed
Lead/Manager

Job Description

Key Skills Required

Master these to land this role

PythonBestseller 🔥
Learn in 56 Hours
DevOpsBestseller 🔥
Learn in 63 Hours
KubernetesAWSTerraform

Want to know if you're a match for this job?

Calculate My Match Score

The Challenge

EverOps is looking for a Lead DevOps Engineer with deep AWS infrastructure experience and unusually strong networking expertise to support a complex cloud migration and modernization initiative within a high-transaction payments environment.

You’ll be joining an active migration already in motion, where timelines are compressed, dependencies are not always fully documented, and the environment spans legacy infrastructure, AWS, application platforms, databases, networking, and production operations.

This role requires someone who can get productive quickly, work through ambiguity, and independently turn broad objectives into executable technical work.

The Mission

As a Lead DevOps Engineer, you will join our U.S.-Based Virtual Operating Center and embed directly with a customer engineering team.

Your immediate priority will be providing hands-on engineering leadership and execution for an active migration from legacy hosted infrastructure into AWS. You’ll troubleshoot migration blockers, assess network and infrastructure dependencies, build cloud infrastructure, and help move production workloads safely.

As the immediate migration effort stabilizes, your focus will expand into modernizing legacy CentOS workloads, building the supporting deployment and observability platform, and implementing AWS-based disaster recovery for a business-critical payments platform.

This is a highly autonomous role. You will be expected to identify what needs to happen, define technical deliverables, communicate risks and dependencies clearly, and drive work through completion without requiring step-by-step direction.

What You’ll Do

  • Cloud Migration: Provide hands-on engineering support for the migration of production workloads from legacy hosted infrastructure into AWS.

  • Network Engineering: Diagnose and resolve complex connectivity, routing, DNS, firewall, VPN, load-balancing, security-group, and hybrid-networking issues impacting migrations and production systems.

  • Migration Planning: Assess undocumented or partially documented environments, identify dependencies and blockers, and translate findings into practical migration plans and technical workstreams.

  • AWS Infrastructure: Design, build, troubleshoot, and improve production AWS environments using modern infrastructure-as-code practices.

  • Legacy Modernization: Help retire legacy CentOS 7 systems and migrate applications onto a modern, supportable platform.

  • Platform Engineering: Implement and improve CI/CD, monitoring, observability, secrets management, configuration management, and deployment automation.

  • Disaster Recovery: Design and implement AWS-based disaster recovery capabilities, including infrastructure, replication dependencies, recovery procedures, and operational runbooks.

  • Production Cutovers: Support migration rehearsals, rollback planning, production cutovers, validation, and post-migration stabilization.

  • Technical Ownership: Independently define and execute technical deliverables, surface risks early, and drive issues to resolution.

  • Documentation: Produce useful architecture documentation, migration plans, operational runbooks, dependency maps, and knowledge-transfer materials.

  • Technical Leadership: Serve as a senior technical partner to customer engineers and EverOps team members, providing direction when ambiguity or complex infrastructure decisions arise.

You Have

  • Experience: 7+ years of professional experience in DevOps, Cloud Engineering, SRE, Infrastructure Engineering, or a related discipline, with significant production AWS experience.

  • AWS Expertise: Deep hands-on experience designing, operating, troubleshooting, and migrating production workloads in AWS.

  • Networking Depth: Strong knowledge of TCP/IP, routing, subnetting, DNS, NAT, firewalls, VPNs, proxies, load balancers, security groups, network ACLs, and AWS networking services.

  • AWS Networking: Production experience with VPC architecture, Transit Gateway, Route 53, ALB/NLB, PrivateLink/VPC endpoints, VPN connectivity, and multi-account or hybrid-network environments.

  • Migration Experience: Proven experience executing data-center, hosted-infrastructure, lift-and-shift, re-platforming, or cloud-modernization migrations involving live production workloads.

  • Infrastructure as Code: Advanced proficiency with Terraform and experience managing production infrastructure through version-controlled IaC.

  • Linux: Strong Linux systems administration and troubleshooting skills, including experience with legacy environments and operating-system modernization.

  • Containers: Production experience with Docker and container orchestration platforms such as ECS, EKS, or Kubernetes.

  • CI/CD: Experience designing and operating modern CI/CD pipelines using tools such as GitHub Actions, Jenkins, Argo CD, or similar platforms.

  • Observability: Experience implementing and troubleshooting monitoring, logging, metrics, and alerting platforms such as Datadog, Prometheus, Grafana, or comparable tooling.

  • Automation: Strong scripting ability using Python, Bash, or similar languages to automate infrastructure and operational workflows.

  • Production Operations: Experience supporting business-critical production environments where uptime, change control, and careful migration planning matter.

  • Autonomy: Demonstrated ability to enter an unfamiliar environment, independently identify priorities, define a path forward, and execute with minimal supervision.

  • Communication: Ability to clearly explain technical risks, dependencies, tradeoffs, and recommendations to both engineers and technical leadership.

Extra Awesome

  • Fintech / Payments: Experience working in payments, financial services, banking, or another highly regulated, high-transaction environment.

  • Rackspace / Hosted Infrastructure: Experience migrating workloads out of Rackspace or similar managed hosting / colocation environments.

  • Migration Leadership: Experience owning end-to-end infrastructure migrations, including discovery, dependency mapping, rehearsal, cutover, rollback, and stabilization.

  • Disaster Recovery: Hands-on experience designing and implementing AWS DR environments, replication strategies, recovery runbooks, and RTO/RPO objectives.

  • Database Infrastructure: Familiarity with infrastructure supporting Aurora/MySQL, PostgreSQL, Cassandra, or other distributed data platforms.

  • Security: Experience operating within environments involving PCI, SOC 2, or similar security and compliance requirements.

  • GitOps: Experience with GitOps and pull-request-driven infrastructure workflows using tools such as Atlantis, Terraform Cloud/Enterprise, Scalr, or Argo CD.

  • Platform Engineering: Experience building internal platforms that standardize deployment, observability, secrets management, and infrastructure consumption.

  • Certifications: AWS Certified Solutions Architect – Professional, AWS Certified Advanced Networking – Specialty, CKA, or similar advanced certifications.

How would you rate this job post?

See what other professionals think about this role.

banner

EverOps (operating via everops.com) is a premier Embedded Service Provider and strategic technology partner engineered to help global enterprise software companies accelerate software delivery, reduce operating risk, and optimize cloud spend. Founded in 2012 and headquartered in San Francisco, California, the enterprise fundamentally disrupts traditional staff augmentation and consulting models by deploying embedded 'TechPods'—elite engineering teams that integrate directly into client environments with proven playbooks and AI-native practices. Moving far beyond advisory roles, EverOps natively unifies Cloud, CI/CD, Observability, Site Reliability Engineering (SRE), Security, and FinOps to deliver guaranteed outcomes on complex, mission-critical initiatives like re-platforming and global network modernization. The company partners closely with industry-leading platform providers, holding prestigious AWS Select Tier and Datadog Advanced partner certifications, to ensure deep architectural expertise and direct vendor support. Under the hood, their execution methodology leverages forward-deployed engineers who co-own high-stakes infrastructure challenges, utilizing infrastructure-as-code (IaC), GitOps delivery models, and comprehensive observability integrations to eliminate technical debt and significantly improve mean time to resolution (MTTR). Capturing massive market acceleration, having supported multiple IPOs and creating billions in combined client market value, EverOps remains a definitive cornerstone for organizations seeking to turn complex operational bottlenecks into highly scalable, automated, and secure digital foundations.

Safety First

  • Never pay for a job application.
  • Do not share sensitive bank info.
  • Verify the client before starting work.
Learn More
Lead DevOps Engineer (Remote) at EverOps