Job Overview
We are seeking an experienced DevOps Engineer to join a growing DevOps organization focused on reliable, secure and scalable cloud platforms. The person in this role will operate, automate, monitor and continuously improve AWS-based infrastructure and deployment pipelines to ensure long-term platform stability and strong application operations.
Responsibilities
- Operate, monitor and support production applications and their AWS cloud infrastructure.
- Detect, troubleshoot and resolve incidents, errors and operational issues across applications and infrastructure.
- Manage security findings and implement patching and update activities as required.
- Develop and maintain Infrastructure-as-Code using Terraform or OpenTofu.
- Administer and optimize Kubernetes clusters and related configuration.
- Support deployment and operation of containerized applications using Docker, Helm and GitOps practices (e.g., Flux).
- Analyze and resolve deployment and release management issues.
- Manage AWS IAM roles, permissions and cloud security configurations.
- Support integration and operation of PostgreSQL databases hosted on AWS RDS.
- Manage secrets, encryption and key management using AWS Secrets Manager and AWS KMS.
- Operate and enhance monitoring and logging based on Grafana and OpenSearch.
- Support and maintain Kafka infrastructure using Amazon MSK Serverless.
- Optimize Kubernetes node scaling and resource management with Karpenter.
- Maintain automated dependency management and update processes with Renovate.
- Develop and support AWS Lambda functions using Python.
- Investigate infrastructure, deployment, database, IAM and integration issues and propose technical solutions.
- Create and maintain operational documentation, runbooks and support procedures.
- Drive standardization, automation and continuous improvement of operational processes to increase platform reliability and availability.
Qualifications
- Minimum 3 years of operational experience running production IT applications in cloud environments.
- Minimum 3 years of experience in configuration and administration of AWS cloud infrastructure.
- Minimum 3 years of experience with Kubernetes for provisioning and managing containerized applications.
- Minimum 3 years of Infrastructure-as-Code experience with Terraform or OpenTofu.
- Experience operating business-critical applications with established incident management processes (up to 3 years).
- Fluent German (C1) is required.
- Practical experience with Docker, Helm, GitOps (Flux), Grafana, OpenSearch, MSK Serverless, Karpenter, Renovate and AWS Lambda with Python is strongly desirable.
Benefits
- Remote working.
- Unique TEAL culture, relationship- and respect-driven community, non-corporate atmosphere.
- Agile approach and no bureaucracy.
- Outstanding integration trips to various places in Europe.
- Activities to support your well-being and health.
- Luxmed Gold Extended medical care and Multisport Plus benefit.