We are looking for a Senior Infrastructure Engineer to design, build, automate, and operate infrastructure across our on-prem data centers and GCP environments. This role suits an engineer who came up through hands-on operations (racking hardware, running networks, bare metal OS installs, keeping systems online at 2am) and has since built deep expertise in automation, infrastructure as code, and public cloud. You'll set technical direction for infrastructure, mentor engineers, and own the systems that scale our global operations.
Responsibilities
-
Design and operate large-scale, multi data center infrastructure supporting international operations, spanning on-prem hardware, GCP, and AWS
-
Lead the automation of infrastructure deployment and configuration management, driving migration from legacy tooling to modern Infrastructure as Code solutions such as Terraform, Puppet, and Ansible
-
Design and maintain traffic management, load balancing, and DNS systems at scale
-
Build and operate monitoring and observability systems across both on-prem and cloud environments
-
Develop tooling to support distributed, auditable systems administration
-
Write and maintain process, policy, and procedural documentation for the infrastructure organization
Requirements
-
A minimum of 3 years of relevant experience in infrastructure or systems engineering, including direct operational ownership of production systems
-
Experience with all aspects of remote management of physical hardware infrastructure colocated at hosting facilities
-
Deep Linux systems administration experience, such as Red Hat or Debian, at scale
-
Hands-on experience with core networking, including Cisco hardware, DNS, load balancing, and traffic management
-
Experience with enterprise storage and backup systems
-
Proven track record automating infrastructure using tools such as Ansible, Puppet, or similar technologies
-
Proficiency in Python or a similar language for infrastructure tooling and automation
-
Production experience with GCP, including Compute Engine, networking, storage, and IAM
-
Demonstrated ability to lead infrastructure projects and mentor other engineers
-
Excellent English proficiency (B2 level or higher)
Nice to have
-
Experience migrating on-prem workloads to public cloud
-
Experience with distributed monitoring and logging stacks, such as Grafana or Prometheus
-
Experience with container orchestration, such as Kubernetes or GKE
-
Experience mentoring and leading junior engineers
We offer
-
International projects with top brands
-
Work with global teams of highly skilled, diverse peers
-
Healthcare benefits
-
Employee financial programs
-
Paid time off and sick leave
-
Upskilling, reskilling and certification courses
-
Unlimited access to the LinkedIn Learning library and 22,000+ courses
-
Global career opportunities
-
Volunteer and community involvement opportunities
-
EPAM Employee Groups
-
Award-winning culture recognized by Glassdoor, Newsweek and LinkedIn
EPAM is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, disability, protected veteran status, or any other characteristic protected by applicable law.