RemoteFull timeMid levelPosted today
Apply with JobAssistAbout the role
- We are looking for an experienced Observability Infrastructure Engineer to join our Platform Engineering organization
- You will be part of the team responsible for building and running Observability pillars on premise and on Kubernetes
- Our systems collect, process, and store the logs, metrics, and traces that allow hundreds of product teams to monitor their services in real time
- You will work in a large-scale environment where we manage petabytes of data and thousands of servers
- We are currently in the middle of a major transformation: focusing on automation of operations and enabling self service for our users
- Build the next generation of our platform: Design and implement the future architecture of our logging and metrics systems. You will play a key role in redesigning our infrastructure to support new global regions, ensuring data isolation and regulatory compliance in different geographies, and more
- Own infrastructure operations: You will take full ownership of our hybrid infrastructure, managing the lifecycle of over 1,500 servers across both bare-metal and Kubernetes environments
- Automate to reduce toil: You will write code in Go or Python to eliminate manual operational tasks. Your goal is to build self-healing systems that do not require manual intervention during the night. You will improve our CI pipelines to ensure that changes to our clusters are safe, predictable, and automated
- Optimize for scale and performance: You will dive deep into performance bottlenecks within our distributed tracing and logging pipelines. We deal with high-volume data streams that can overwhelm standard configurations. You will tune our Elasticsearch clusters, optimize Prometheus and VictoriaMetrics storage, and ensure our OpenTelemetry implementation can handle peak traffic without missing a beat
- Reliability and Engineering: You will participate in on-call rotations, but your primary focus will be engineering solutions that stop alerts from firing in the first place. You will help us upgrade our stack to the latest versions and ensure our platform remains secure and performant. You will improve the self-service experience by implementing automated guardrails and quota management to prevent noisy tenants from destabilizing the platform, while designing safer API access patterns for our users
Benefits
- Global exchange program
- Weekly happy hour
- Delicious healthy lunches
- Phantom share package
- Yearly trip to Amsterdam
- Paid holidays
- Work from home opportunities- This is a role for a builder and a problem solver who enjoys deep technical troubleshooting across distributed systems and then turns recurring issues into automated, repeatable solutions
- Observability Stack Expertise: You have hands-on experience operating core telemetry data stores at scale e.g. Elasticsearch/Opensearch/VictoriaLogs/Clickhouse for logging, Prometheus/ VictoriaMetrics for metrics and Grafana Tempo for distributed tracing
- Production Kubernetes Experience: Proven hands-on experience operating, and troubleshooting production workloads on Kubernetes (on-prem and/or cloud), including strong day-to-day use of kubectl and Kubernetes primitives (e.g. Namespaces, Pods, Deployments/StatefulSets, Services, Ingress, ConfigMaps/Secrets)
- 10+ years of experience in the observability domain or in a relevant platform/infrastructure domain
- Software Engineering Mindset: You are proficient in Go or Python and do not just write scripts; you build tools and automation platforms that treat infrastructure as code
- Linux Experience: You understand the operating system at a kernel level and can debug complex networking, file system, and performance issues on both bare metal and virtualized hardware
- Experience with large scale, multi tenant isolation and quota or cost governance approaches for telemetry platforms.
- Familiarity with regulated environments where security, audibility, and data handling requirements shape platform design decisions.
- Studies show that women and members of underrepresented communities apply for jobs only if they meet 100% of the qualifications. Does this sound like you? If so, Adyen encourages you to reconsider and apply. We look forward to your application!- Ensuring a smooth and enjoyable candidate experience is critical for us. We aim to get back to you regarding your application within 10 business days. Our interview process tends to take about 4 weeks to complete but may fluctuate depending on the role
Millions of jobs, with real people getting hired every day
20,000+
New jobs added daily7,000,000+
Verified job listings500,000+
Tailored applications submittedFAQ
Questions, answered
Click "Apply with JobAssist" – we tailor your resume and application to this role and submit it for your approval.
Yes. This role at Adyen was screened before publishing – we confirmed the employer before listing it.
The employer didn't disclose a salary range for this listing. JobAssist shows pay whenever it's available.
This position can be done from anywhere, with no in-office requirement.
Yes – every application is tailored from your profile and this job's requirements, and you can review and edit before it's sent.
