Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме
описание
Adyen provides payments, data, and financial products in a single solution for customers such as Facebook, Uber, H&M, and Microsoft. It operates a financial technology platform that helps businesses achieve their ambitions through integrated financial services.
*Instagram и Facebook принадлежат компании Meta Platforms Inc., деятельность которой признана экстремистской и запрещена на территории РФ
задачи
Design and implement the future architecture of logging and metrics systems;
Redesign infrastructure to support new global regions, data isolation, and regulatory compliance;
Own the hybrid infrastructure and manage the lifecycle of over 1,500 servers across bare-metal and Kubernetes environments;
Write Go or Python code to eliminate manual operational tasks;
Build self-healing systems and improve CI pipelines for safe, predictable, and automated cluster changes;
Investigate performance bottlenecks in distributed tracing and logging pipelines;
Tune Elasticsearch clusters and optimize Prometheus and VictoriaMetrics storage;
Ensure the OpenTelemetry implementation handles peak traffic reliably;
Participate in on-call rotations;
Upgrade the technology stack and keep the platform secure and performant;
Implement automated guardrails and quota management;
Design safer API access patterns for users.
требования
10+ Years of experience in observability or a relevant platform/infrastructure domain;
Hands-on experience operating core telemetry data stores at scale, including Elasticsearch, Opensearch, VictoriaLogs, ClickHouse, Prometheus, VictoriaMetrics, and Grafana Tempo;
Kernel-level Linux knowledge and the ability to debug complex networking, file system, and performance issues on bare metal and virtualized hardware;
Proven experience operating and troubleshooting production Kubernetes workloads on-premises and/or in the cloud;
Strong day-to-day experience with kubectl and Kubernetes primitives, including Namespaces, Pods, Deployments/StatefulSets, Services, Ingress, ConfigMaps, and Secrets;
Proficiency in Go or Python and experience building infrastructure-as-code tools and automation platforms;
Nice to have: large-scale multi-tenant isolation, quota or cost governance for telemetry platforms, regulated environments where security, auditability, and data handling requirements shape platform design.
условия
The role is based out of the Amsterdam office;
Office-first work environment with in-person collaboration; remote-only roles are not offered.