Senior Platform Engineer - Observability

Adyen
Amsterdam
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 10+ yearsSkills: ["Troubleshooting","Problem-solving","Automation","Reliability engineering"]

Build and run observability pillars for hundreds of product teams, designing next-generation logging and metrics infrastructure for hybrid environments and on Kubernetes. Own the lifecycle of 1,500+ servers across bare metal and Kubernetes, automate operations with Go or Python, and improve CI pipelines for safe cluster changes. Optimize distributed tracing and logging at high scale by tuning telemetry stores and OpenTelemetry performance, while delivering reliability through self-healing systems and guardrails.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Adyen
Adyen
6 months ago

Senior Platform Engineer - Observability

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 19 hours agoStatus: Live

Job Summary

Build and run observability pillars for hundreds of product teams, designing next-generation logging and metrics infrastructure for hybrid environments and on Kubernetes. Own the lifecycle of 1,500+ servers across bare metal and Kubernetes, automate operations with Go or Python, and improve CI pipelines for safe cluster changes. Optimize distributed tracing and logging at high scale by tuning telemetry stores and OpenTelemetry performance, while delivering reliability through self-healing systems and guardrails.
Location: Amsterdam
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design and implement the future architecture of logging and metrics systems to support new global regions, data isolation, and regulatory compliance.
  • •Own observability infrastructure operations across hybrid environments, managing the lifecycle of 1,500+ servers on bare metal and Kubernetes.
  • •Automate operational tasks in Go or Python and build self-healing systems that reduce or eliminate manual night-time intervention.
  • •Improve CI pipelines so cluster changes are safe, predictable, and automated.
  • •Optimize performance and scale for distributed tracing and logging, including tuning Elasticsearch clusters and storage (Prometheus/VictoriaMetrics) and ensuring OpenTelemetry handles peak traffic without loss.

Key Requirements

  • •10+ years of experience in the observability domain or a relevant platform/infrastructure domain.
  • •Hands-on expertise operating core telemetry data stores at scale (e.g., Elasticsearch/OpenSearch/VictoriaLogs/ClickHouse for logs; Prometheus/VictoriaMetrics for metrics; Grafana Tempo for tracing).
  • •Strong Linux experience, able to debug complex networking, file system, and performance issues on bare metal and virtualized hardware.
  • •Proven production Kubernetes experience troubleshooting workloads (on-prem and/or cloud), including day-to-day use of kubectl and Kubernetes primitives.
  • •Proficiency in Go or Python with an infrastructure-as-code mindset to build automation tools and platforms.
Experience:10+ yearsObservabilityDistributed systemsHybrid infrastructureKubernetesTelemetry platformsMulti-tenant environmentsRegulated environments
Skills:TroubleshootingProblem-solvingAutomationReliability engineering
Tech Stack:GoPythonElasticsearchOpenSearchVictoriaLogsClickHousePrometheusVictoriaMetricsGrafana TempoOpenTelemetryKubernetesKubectlLinux

Company Brief

Adyen
Global payments company providing a unified platform for accepting payments, risk management, and financial services to merchants and enterprises across online, mobile, and in-store channels.
Industry: Payments
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Amsterdam, Netherlands
Founded: 2006
WebsiteLinkedIn