---
title: Senior Infrastructure Engineer at EPAM Systems
description: We are seeking a Senior Infrastructure Engineer with deep experience in cloud infrastructure, Linux internals, large-scale distributed systems, workload isolation, and a strong security mindset to joi
---

# Senior Infrastructure Engineer

**Company:** EPAM Systems  
**Location:** Georgia  
**Posted:** 2026-09-03  
**Apply by:** 2026-10-18

[Apply / View original posting](https://www.linkedin.com/jobs/view/4461900916)

## Job description

We are seeking a Senior Infrastructure Engineer with deep experience in cloud infrastructure, Linux internals, large-scale distributed systems, workload isolation, and a strong security mindset to join our team and help design, build, and operate a global, resilient platform. To discover more about Cloud practice at EPAM Georgia, visit this page. Experience the freedom of remote work from anywhere in Georgia, whether from the comfort of your home, our modern offices in Tbilisi and Batumi or a coworking space in Kutaisi. Responsibilities Design, build, and maintain a global, distributed, and resilient cloud infrastructure Collaborate with infrastructure and product engineering teams to plan and deliver complex platform initiatives Participate in architecture reviews, incident response, and performance analysis to ensure system reliability Manage and provision AWS infrastructure using Terraform and Kubernetes Write and maintain Kubernetes manifests and deployment configurations for critical workloads, including pod security contexts, anti-affinity rules, network policies, autoscaling, and health probes Drive Production Readiness Reviews (PRR) for all new services, covering security, HA, performance, and observability gates Design and operate multi-layer workload isolation using Linux kernel primitives: namespaces (pid, net, mnt, user, uts, ipc), cgroups, seccomp profiles, and capabilities, as the baseline security boundary Evaluate and operate gVisor and Firecracker for workloads requiring hard tenant boundaries and near-native performance Design and maintain Grafana dashboards Manage Prometheus and VictoriaMetrics pipelines; define and tune P1/P2/P3 alert thresholds with runbooks Contribute to and extend the internal k6-based load testing framework (load-testing-framework / library/k6/webhooks) Design load scenarios using constant-arrival-rate profiles; instrument custom metrics (job_succeeded_count, job_failed_count) tagged by testid for Grafana correlation; stream test metrics to Prometheus via remote write; generate and publish HTML reports to file storage after each run Requirements 5+ years of experience in infrastructure, SRE, or platform engineering roles Expertise in distributed systems and cloud-native architectures Understanding of Linux internals: namespaces, cgroups, seccomp, capabilities, and system-level performance tuning Experience operating infrastructure on AWS at scale Proficiency in Terraform and Kubernetes, including security hardening of manifests Experience designing and running load tests (k6, Gatling, Locust, or similar) Understanding of network security and cloud security best practices Excellent analytical, troubleshooting, and communication skills Proficiency in English at a B2+ level Nice to have Skills in Golang/Python for building internal tooling Familiarity with Kafka, Redis, ClickHouse, or PostgreSQL Familiarity with observability tools such as Grafana, Prometheus, or VictoriaMetrics Knowledge of encryption key hierarchies (CMK/DEK/KMS patterns) and HashiCorp Vault Hands-on production experience with gVisor or Firecracker Prior work extending or maintaining an internal testing framework Contributions to or maintenance of open-source infrastructure projects We offer We connect like-minded people Delivering innovative solutions to industry leaders, making a global impact Enjoyable working environment, whether it is the vibrant office or the comfort of your own home Opportunity to work abroad for up to two months per year Relocation opportunities within our offices in 55+ countries Corporate and social events We invest in your growth Leadership development, career advising, soft skills and well-being programs Certifications, including GCP, Azure and AWS Unlimited access to EPAM's internal learning database Free English classes with certified teachers We cover it all Participation in the Employee Stock Purchase Plan Monetary bonuses for engaging in the referral program Comprehensive medical & family care package Five trust days per year (sick leave without a medical certificate) Benefits package (sports activities, a variety of stores and services) EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.

---

🔔 [Monitor similar jobs on Gurify](https://gurify.com/?utm_source=techgeo&utm_medium=jobboard&utm_campaign=monitor_similar&role=Senior+Infrastructure+Engineer) — get alerted when matching roles are posted in Georgia.

```json
{"@context":"https://schema.org/","@type":"JobPosting","title":"Senior Infrastructure Engineer","description":"<p>We are seeking a Senior Infrastructure Engineer with deep experience in cloud infrastructure, Linux internals, large-scale distributed systems, workload isolation, and a strong security mindset to join our team and help design, build, and operate a global, resilient platform. To discover more about Cloud practice at EPAM Georgia, visit this page. Experience the freedom of remote work from anywhere in Georgia, whether from the comfort of your home, our modern offices in Tbilisi and Batumi or a coworking space in Kutaisi. Responsibilities Design, build, and maintain a global, distributed, and resilient cloud infrastructure Collaborate with infrastructure and product engineering teams to plan and deliver complex platform initiatives Participate in architecture reviews, incident response, and performance analysis to ensure system reliability Manage and provision AWS infrastructure using Terraform and Kubernetes Write and maintain Kubernetes manifests and deployment configurations for critical workloads, including pod security contexts, anti-affinity rules, network policies, autoscaling, and health probes Drive Production Readiness Reviews (PRR) for all new services, covering security, HA, performance, and observability gates Design and operate multi-layer workload isolation using Linux kernel primitives: namespaces (pid, net, mnt, user, uts, ipc), cgroups, seccomp profiles, and capabilities, as the baseline security boundary Evaluate and operate gVisor and Firecracker for workloads requiring hard tenant boundaries and near-native performance Design and maintain Grafana dashboards Manage Prometheus and VictoriaMetrics pipelines; define and tune P1/P2/P3 alert thresholds with runbooks Contribute to and extend the internal k6-based load testing framework (load-testing-framework / library/k6/webhooks) Design load scenarios using constant-arrival-rate profiles; instrument custom metrics (job_succeeded_count, job_failed_count) tagged by testid for Grafana correlation; stream test metrics to Prometheus via remote write; generate and publish HTML reports to file storage after each run Requirements 5+ years of experience in infrastructure, SRE, or platform engineering roles Expertise in distributed systems and cloud-native architectures Understanding of Linux internals: namespaces, cgroups, seccomp, capabilities, and system-level performance tuning Experience operating infrastructure on AWS at scale Proficiency in Terraform and Kubernetes, including security hardening of manifests Experience designing and running load tests (k6, Gatling, Locust, or similar) Understanding of network security and cloud security best practices Excellent analytical, troubleshooting, and communication skills Proficiency in English at a B2+ level Nice to have Skills in Golang/Python for building internal tooling Familiarity with Kafka, Redis, ClickHouse, or PostgreSQL Familiarity with observability tools such as Grafana, Prometheus, or VictoriaMetrics Knowledge of encryption key hierarchies (CMK/DEK/KMS patterns) and HashiCorp Vault Hands-on production experience with gVisor or Firecracker Prior work extending or maintaining an internal testing framework Contributions to or maintenance of open-source infrastructure projects We offer We connect like-minded people Delivering innovative solutions to industry leaders, making a global impact Enjoyable working environment, whether it is the vibrant office or the comfort of your own home Opportunity to work abroad for up to two months per year Relocation opportunities within our offices in 55+ countries Corporate and social events We invest in your growth Leadership development, career advising, soft skills and well-being programs Certifications, including GCP, Azure and AWS Unlimited access to EPAM's internal learning database Free English classes with certified teachers We cover it all Participation in the Employee Stock Purchase Plan Monetary bonuses for engaging in the referral program Comprehensive medical & family care package Five trust days per year (sick leave without a medical certificate) Benefits package (sports activities, a variety of stores and services) EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.</p>","identifier":{"@type":"PropertyValue","name":"TechGeo","value":"4461900916"},"url":"https://techgeo.ge/job/senior-infrastructure-engineer-56fc4fc","jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Georgia","addressCountry":"GE"}},"hiringOrganization":{"@type":"Organization","name":"EPAM Systems"},"directApply":false,"datePosted":"2026-09-03","validThrough":"2026-10-18T23:59:59+04:00","jobLocationType":"TELECOMMUTE","applicantLocationRequirements":{"@type":"Country","name":"Georgia"}}
```
