Tech Stack
KubernetesCI/CDTerraformPrometheusGrafanaArgoCDHelmOpenTelemetryTempoAWS EKS
Job Description, Responsibilities & Requirements
About the Position
We are looking for a Senior DevOps Engineer to help develop and maintain reliable cloud infrastructure, improve deployment processes, and ensure the stability of production environments. In this role, you will contribute to the evolution of our infrastructure and support the continuous growth of our products.
Remote | Full-time | DevOps
Responsibilities
- Administration and development of Kubernetes (EKS) in dev / staging / prod.
- GitOps delivery (ArgoCD / Flux), support of Helm charts and deployment strategies.
- Support and development of CI/CD pipelines (GitLab), release models, per-environment alerts.
- Ownership of the observability stack: Prometheus, Grafana, Alertmanager, OpenTelemetry, Tempo, Loki, OpenSearch.
- Incident response and writing RCA for production incidents.
- Infrastructure as Code (Terraform) on AWS: EKS, RDS, ElastiCache, VPC, S3.
- Administration of databases and messaging: PostgreSQL, MongoDB, Redis, NATS.
- Access management (LDAP, Keycloak) and secrets management (Vault).
- AWS cost optimization.
- Creation and maintenance of technical documentation.
Requirements
- Kubernetes in production (AWS EKS) - dev / staging / prod.
- AWS: EKS, RDS, ElastiCache, VPC, IAM, S3, CloudWatch.
- Infrastructure as Code - Terraform.
- GitOps - ArgoCD and/or Flux (WeaveGitOps), Helm charts, per-environment values.
- CI/CD - GitLab CI/CD, self-hosted runners, release models.
- Observability: Prometheus, Grafana, Alertmanager, OpenTelemetry / Tempo (tracing), Loki (logs), OpenSearch.
- Production databases:
- PostgreSQL (AWS RDS).
- MongoDB (Atlas).
- Redis / ElastiCache.
- Message brokers - NATS (nice to have).
- Docker, Linux, networking fundamentals.
- Scripting - Bash / Python.
- Incident response, RCA, working with alerts.
Nice to Have
- HashiCorp Vault.
- Keycloak / LDAP.
- Gateway API / service mesh (migration from Ingress-NGINX).
- CDN - Cloudflare, CloudFront, failover between providers.
- AWS cost optimization - autoscaling, spot, savings plans, MAP tagging.
- BullMQ / Bull Board.
Your Journey with Us
- Step 1: Pre-screen.
- Step 2: Technical interview.
- Step 3: Final interview.
- Step 4: Reference check.
- Step 5: Job Offer!
We Offer
- 28 business days of paid off.
- Flexible hours and the possibility to work remotely.
- Medical insurance and mental health care.
- Compensation for courses, trainings.
- English classes and speaking clubs.
- Internal library, educational events.
- Outstanding corporate parties, teambuildings.
About the Company
For CV / career questions
[email protected]
For all other questions
[email protected]
Apply for a vacancy
Kate Kravchenko
Recruiter
ask a question