Strength. Care. Growth
You will know we are the right place for you, if you are driven by:
- Opportunities to learn and build your career.
- Meaningful work in a stable and fast-paced company.
- Diversity of people, projects, and platforms.
- A supportive, fun, and inspiring place to work.
This job can be performed by all countries within our A1 footprint.
Job Overview:
We are looking for a KubeForge Engineer to join our Kubernetes-focused team. In this role, you will be responsible for operating, maintaining, and optimizing our Kubernetes environments—including OpenShift, RKE2, and vanilla Kubernetes clusters—across the organization. You will play a key role in ensuring high availability, performance, and security of our container orchestration platforms, directly impacting the reliability of services relied upon by the entire A1 Telekom group.
Role Insights:
- Administer and maintain production Kubernetes clusters (OpenShift, RKE2, and vanilla Kubernetes) with a focus on availability and performance.
- Deploy, configure, and manage containerized workloads using Kubernetes primitives (Deployments, StatefulSets, Services, ConfigMaps, Secrets).
- Implement and enforce security best practices, including RBAC, network policies, Pod Security Standards, and certificate management.
- Monitor cluster health using tools such as Prometheus, Grafana, and built-in OpenShift observability stack; respond to and resolve incidents.
- Automate operational tasks and cluster management using Infrastructure-as-Code tools (e.g., Terraform, Ansible).
- Collaborate with development teams to support CI/CD pipelines and containerized application deployments.
- Conduct capacity planning, performance tuning, and regular maintenance activities (patching, upgrades).
- Document cluster architectures, operational procedures, and runbooks.
What Makes You Unique:
- 3+ years of hands-on experience with Kubernetes administration in production environments.
- Deep understanding of Kubernetes architecture, networking (CNI plugins, Ingress, DNS), and resource management.
- Experience with at least one managed Kubernetes platform (OpenShift preferred, EKS/AKS/GKE acceptable).
- Proficiency in Linux system administration and command-line tools
- Familiarity with monitoring, logging, and observability solutions (Prometheus, Grafana, ELK stack).
- Strong problem-solving skills with the ability to work under pressure during incidents.
- Fluency in English language.