Description
Summary:
Seeking a Lead Infrastructure Engineer to design, build, automate, and oversee infrastructure across on-prem and GCP environments, shaping technical direction and guiding engineers.
Highlights:
1. Shape technical direction for infrastructure and guide engineers
2. Lead automation initiatives for infrastructure deployment
3. Design and oversee large-scale, multi data center infrastructure
We are seeking a Lead Infrastructure Engineer to design, build, automate, and oversee infrastructure across our on\-prem data centers and GCP environments. This position is well\-suited to an engineer who built their foundation through hands\-on operations work (racking hardware, running networks, bare metal OS installs, keeping systems online at 2am) and has since developed strong expertise in automation, infrastructure as code, and public cloud platforms. You'll shape technical direction for infrastructure, guide engineers, and take ownership of the systems powering our worldwide operations.
**Responsibilities**
* Design and oversee large\-scale, multi data center infrastructure that supports global operations across on\-prem hardware, GCP, and AWS
* Lead automation initiatives for infrastructure deployment and configuration management, steering the transition from legacy tooling to modern Infrastructure as Code solutions such as Terraform, Puppet, and Ansible
* Design and support traffic management, load balancing, and DNS systems at scale
* Build and manage monitoring and observability systems across both on\-prem and cloud environments
* Develop tooling that supports distributed, auditable systems administration
* Write and update process, policy, and procedural documentation for the infrastructure team
**Requirements**
* A minimum of 5 years of relevant experience in infrastructure or systems engineering, including direct operational responsibility for production systems
* Experience managing all aspects of remote administration for physical hardware infrastructure hosted at colocation facilities
* Substantial Linux systems administration experience, such as Red Hat or Debian, at scale
* Hands\-on experience with core networking, including Cisco hardware, DNS, load balancing, and traffic management
* Experience with enterprise storage and backup systems
* Solid track record automating infrastructure using tools such as Ansible, Puppet, or similar technologies
* Proficient in Python or a similar language for infrastructure tooling and automation
* Production experience with GCP, including Compute Engine, networking, storage, and IAM
* Proven ability to lead infrastructure projects and mentor fellow engineers
* Excellent English proficiency (B2 level or higher)
**Nice to have**
* Experience shifting on\-prem workloads to public cloud platforms
* Experience with distributed monitoring and logging stacks, such as Grafana or Prometheus
* Experience with container orchestration tools, such as Kubernetes or GKE
* Experience mentoring and guiding junior engineers
**We offer**
* International projects with top brands
* Work with global teams of highly skilled, diverse peers
* Healthcare benefits
* Employee financial programs
* Paid time off and sick leave
* Upskilling, reskilling and certification courses
* Unlimited access to the LinkedIn Learning library and 22,000\+ courses
* Global career opportunities
* Volunteer and community involvement opportunities
* EPAM Employee Groups
* Award\-winning culture recognized by Glassdoor, Newsweek and LinkedIn
*EPAM is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, disability, protected veteran status, or any other characteristic protected by applicable law.*