DevOps Engineer
- Not specified
Posted 2 months ago
About the role
Observability Platform Engineer
๐ Location: Hamburg, Hamburg, Germany (Hybrid)
๐ข Industry: IT Services and IT Consulting
๐ผ Work Setting: Hybrid
Are you passionate about building and optimizing modern observability and monitoring systems? This is a great opportunity to work on cutting-edge technologies across distributed systems, cloud platforms, and containerized environments.
As an Observability Platform Engineer, you will design, implement, and manage scalable observability solutions, with a strong focus on distributed tracing and the Grafana ecosystem. You’ll play a key role in ensuring system reliability, performance, and visibility across complex infrastructures.
Key Responsibilities
- Design, develop, and enhance observability platforms with a focus on distributed tracing and monitoring
- Build and maintain system architecture integrating metrics, logs, and traces
- Deploy, configure, and scale monitoring solutions using tools like Grafana, Tempo, Loki, and Prometheus
- Implement and manage tracing protocols such as OpenTelemetry, Jaeger, and Zipkin
- Analyze tracing data using TraceQL to identify performance bottlenecks and system issues
- Optimize platform performance through tuning (e.g., sampling, caching strategies)
- Manage observability solutions in containerized environments (Kubernetes/OpenShift)
- Develop and maintain CI/CD pipelines using tools such as Git, Terraform, Ansible, and Artifactory
- Ensure system security, data integrity, and access control (LDAP, SAML, OAuth)
- Monitor infrastructure health and plan for scalability (CPU, memory, storage)
Required Skills & Experience
- Strong experience with observability tools, particularly within the Grafana ecosystem
- Hands-on experience with container technologies and orchestration tools
- Solid understanding of system architecture and distributed systems
- Experience with CI/CD pipelines and infrastructure as code tools
- Working knowledge of cloud platforms (AWS, Azure, or GCP)
- Proficiency in Linux/Unix system administration
- Ability to analyze system performance and resolve complex issues
Preferred Skills
- Experience with OpenShift and enterprise container platforms
- Strong dashboard creation skills for trend analysis and anomaly detection
- Exposure to large-scale infrastructure design and operations
- Familiarity with security and compliance requirements for enterprise systems