
Senior Infrastructure Engineer
Primary stack
Nice to have's
Job description
About the Position
We are seeking a Senior Infrastructure Engineer to architect, build, automate, and manage infrastructure across our on-prem data centers and GCP environments. This position suits an engineer who built their skills through hands-on operations work (racking hardware, running networks, bare metal OS installs, keeping systems online at 2am) and has since gained deep expertise in automation, infrastructure as code, and public cloud platforms. You'll establish technical direction for infrastructure, support engineers' growth, and take ownership of the systems that drive our global operations.
Responsibilities
- Architect and manage large-scale, multi data center infrastructure supporting global operations across on-prem hardware, GCP, and AWS
- Lead automation efforts for infrastructure deployment and configuration management, guiding the move from legacy tooling to modern Infrastructure as Code solutions such as Terraform, Puppet, and Ansible
- Design and maintain traffic management, load balancing, and DNS systems at scale
- Build and oversee monitoring and observability systems across both on-prem and cloud environments
- Create tooling that supports distributed, auditable systems administration
- Draft and maintain process, policy, and procedural documentation for the infrastructure team
Requirements
- At least 3 years of relevant experience in infrastructure or systems engineering, including direct operational responsibility for production systems
- Experience overseeing all aspects of remote administration for physical hardware infrastructure hosted at colocation facilities
- Strong Linux systems administration experience, such as Red Hat or Debian, at scale
- Practical experience with core networking, including Cisco hardware, DNS, load balancing, and traffic management
- Experience with enterprise storage and backup systems
- Proven history automating infrastructure using tools such as Ansible, Puppet, or similar technologies
- Skilled in Python or a similar language for infrastructure tooling and automation
- Production experience with GCP, including Compute Engine, networking, storage, and IAM
- Demonstrated ability to lead infrastructure projects and mentor fellow engineers
- Excellent English proficiency (B2 level or higher)
Nice to Have
- Experience moving on-prem workloads to public cloud platforms
- Experience with distributed monitoring and logging stacks, such as Grafana or Prometheus
- Experience with container orchestration tools, such as Kubernetes or GKE
- Experience mentoring and supporting junior engineers
We Offer
- Remote work in Brazil and other locations
- Opportunity to work with cutting-edge technologies and infrastructure
About the Company
EPAM Systems is a global software engineering and product development company. We deliver digital platforms and software products for the world’s leading companies.
© EPAM. This job description was sourced from the employer's public career page. TheJob is not the employer — we index the posting and route candidates to the source. All content rights and hiring decisions belong to the employer.
EPAM helps organizations innovate their business processes and rethink the way they manage their businesses so they can remain competitive in this new digital age.