DevOps Architect
Реф. №
27718
Модел на работа
На място
Месторабота / Населено място
гр. София
Публикувана на:
22 юни 2026
Отговорности
- Own the strategic development of services portfolio for the DevOps/Platform Engineering organization
- Design and maintain multi-region, production-grade cloud infrastructure on AWS (primary), using Infrastructure as Code. Facilitate architecture design and vision on self-service, lifecycle and day2 operations workflows on company scale
- Architect and operate Kubernetes platforms (EKS, self-managed) including cluster lifecycle, networking, autoscaling, high-availability and durability
- Own the GitOps delivery pipeline end-to-end: ArgoCD for continuous deployment, image promotion strategies, Helm chart management, and environment orchestration. Design promotion strategies spanning hundreds of platform instances
- Build and evolve the observability stack—Grafana, Loki, Mimir, Tempo (LGTM)—including multi-tenant log routing, cross-region forwarding, and cost optimization
- Design networking architectures: VPCs, Transit Gateway, Direct Connect, PrivateLink, DNS (Route 53), and API gateway topologies
- Own the vision for CI/CD pipelines that support rapid, safe delivery across multiple teams and repositories.
- Drive cost optimization initiatives across compute, storage, networking (cross-AZ/cross-region traffic), and observability infrastructure
- Define platform engineering standards: golden paths, self-service capabilities, and internal developer platform (IDP) patterns that accelerate team velocity
- Evaluate, prototype, and adopt new technologies and architectural patterns. Document decisions through ADRs, comparison matrices, and risk analyses
- Participate in incident response, root cause analysis, and reliability improvements for production systems when advanced knowledge of the domain is required
- Steer the organizational efforts in the architecture strategy applying systematic thinking and soft skills
Изисквания
- 5+ years of hands-on experience in cloud infrastructure, platform engineering, or DevOps/SRE roles.
- Deep expertise in AWS services and designs
- Strong Kubernetes skills: cluster design, RBAC, network policies, ingress/gateway API, Helm, operators, and troubleshooting
- Proficiency with Infrastructure as Code (Terraform) including module design, state management, and TACOS automation (Atlantis, Digger, or similar)
- Proficiency with GitOps practices and tools, particularly ArgoCD
- Solid understanding of networking: DNS, load balancing (NLB/ALB), TLS termination, service mesh concepts, and cross-region connectivity
- Solid understanding of production operations processes and facilitation of automation for reduction of toil, disaster recovery and data migration strategies
- Hands-on experience with at least one modern observability stack (Prometheus, Grafana, Loki, Tempo, or equivalent)
- Strong scripting and automation skills (Bash, Python, Javascript, Groovy or Go)
- Good expertice in security principles and practices and Security in System Design
- Excellent documentation and communication skills; ability to produce architecture decision records, risk analyses, and technical comparisons
- Should be able to form coherent and understandable designs and be able to supervice their implementation and delivery to production
Nice to have:
- Familiarity with OpenTelemetry Collector configuration: pipelines, processors, routing connectors, and multi-tenant architectures
- Knowledge of streaming and real-time infrastructure: WebRTC, RTSP
- Experience with globally distributed database architectures at scale: cross-region replication strategies, DR
- Familiarity with API gateways and Kubernetes Gateway API spec evaluation and implementation
- Knowledge of Software Development paradigms, DDD, SOLID practices, Design Paterns and good understanding of Microservice Patterns
- Exposure to secrets management patterns: External Secrets Operator, Vault, CSI Secret Store Driver
Професионална сфера
ИТ - Разработка / поддръжка на хардуер, ИТ - Разработка / поддръжка на софтуер