Job Description
A leading global education technology company is seeking a Principal DevOps Engineer to provide technical leadership within its DevOps and platform engineering function. The role focuses on building reliable, scalable, secure, and developer-friendly infrastructure across AWS serverless and Kubernetes environments.
Location: Krakow, Katowice, Gdansk, Poznan, Poland
Employment Type: Full-time
Work Arrangement: Remote
Key Responsibilities
- Lead the technical design and implementation of platform capabilities across AWS serverless workloads and Kubernetes/EKS environments.
- Partner with architects and senior technical leaders to translate platform strategy into practical, scalable solutions.
- Break down complex architectural initiatives into clear, executable milestones and delivery plans.
- Develop reusable infrastructure patterns, modules, and engineering standards.
- Drive improvements that reduce incident frequency, severity, and operational workload.
- Strengthen observability practices across logging, metrics, tracing, monitoring, and alerting.
- Improve CI/CD automation, deployment workflows, change management, and incident response practices.
- Contribute to Kubernetes cluster architecture, networking, security, and platform modernization.
- Improve cost visibility, resource tagging, logging pipelines, and infrastructure guardrails.
- Establish clear deployment and monitoring pathways that improve the developer experience.
- Support the adoption of DORA metrics and reliability-focused engineering practices.
- Mentor senior engineers, lead technical discussions and design reviews, and communicate technical risks and trade-offs clearly.
- Support the responsible adoption of AI-assisted development tools within engineering workflows.
Requirements
- 8–12+ years of experience in infrastructure, SRE, DevOps, or platform engineering.
- Proven experience operating serverless and Kubernetes environments at scale.
- Strong AWS expertise, including Lambda, API Gateway, Step Functions, DynamoDB, S3, IAM, and VPC.
- Strong Kubernetes/EKS experience, particularly in networking, observability, and production operations.
- Advanced Infrastructure as Code experience, preferably with Terraform.
- Strong experience designing and managing CI/CD pipelines.
- Hands-on experience with observability technologies such as OpenTelemetry, OpenSearch, Prometheus, Grafana, and CloudWatch.
- Strong understanding of distributed systems, reliability engineering, and production infrastructure.
- Demonstrated ability to improve system reliability and reduce operational toil.
- Experience collaborating with multiple engineering teams and senior technical stakeholders.
- Strong technical leadership, communication, problem-solving, and mentoring skills.
Preferred Qualifications
- Experience implementing and using DORA metrics to improve engineering performance.
- Experience supporting SRE operating models and incident reduction initiatives.
- Experience improving developer onboarding and platform user experience.
- Experience developing engineering tools using Python, Node.js, or similar scripting languages.
- Experience evaluating and implementing AI-assisted development tools responsibly.
Key Competencies
- Resourcefulness
- Collaboration & Influencing
- Quality Focus
