
Senior Infrastructure Engineer
Primary stack
Nice to have's
Job description
About the Position
We are seeking a Senior Infrastructure Engineer to design, build, automate, and operate infrastructure across our on-prem data centers and GCP environments. This role suits an engineer who came up through hands-on operations (racking hardware, running networks, bare metal OS installs, keeping systems online at 2am) and has since built deep expertise in automation, infrastructure as code, and public cloud. You'll set technical direction for infrastructure, mentor engineers, and own the systems that scale our global operations.
Responsibilities
- Design and operate large-scale, multi data center infrastructure supporting international operations, spanning on-prem hardware, GCP, and AWS
- Lead the automation of infrastructure deployment and configuration management, driving migration from legacy tooling to modern Infrastructure as Code solutions such as Terraform, Puppet, and Ansible
- Design and maintain traffic management, load balancing, and DNS systems at scale
- Build and operate monitoring and observability systems across both on-prem and cloud environments
- Develop tooling to support distributed, auditable systems administration
- Write and maintain process, policy, and procedural documentation for the infrastructure organization
Requirements
- A minimum of 3 years of relevant experience in infrastructure or systems engineering, including direct operational ownership of production systems
- Experience with all aspects of remote management of physical hardware infrastructure colocated at hosting facilities
- Deep Linux systems administration experience, such as Red Hat or Debian, at scale
- Hands-on experience with core networking, including Cisco hardware, DNS, load balancing, and traffic management
- Experience with enterprise storage and backup systems
- Proven track record automating infrastructure using tools such as Ansible, Puppet, or similar technologies
- Proficiency in Python or a similar language for infrastructure tooling and automation
- Production experience with GCP, including Compute Engine, networking, storage, and IAM
- Demonstrated ability to lead infrastructure projects and mentor other engineers
- Excellent English proficiency (B2 level or higher)
Nice to Have
- Experience migrating on-prem workloads to public cloud
- Experience with distributed monitoring and logging stacks, such as Grafana or Prometheus
- Experience with container orchestration, such as Kubernetes or GKE
- Experience mentoring and leading junior engineers
We Offer
- Opportunity to work remotely in Brazil, Portugal, or Spain
- Competitive contractor terms
About the Company
EPAM Systems is a global software engineering and product development company, delivering digital transformation and technology innovation across industries.
© EPAM. This job description was sourced from the employer's public career page. TheJob is not the employer — we index the posting and route candidates to the source. All content rights and hiring decisions belong to the employer.
EPAM helps organizations innovate their business processes and rethink the way they manage their businesses so they can remain competitive in this new digital age.