Still hiring?
14 checks1 red flag
Browse
All Tech JobsThe whole board, newest first.Roles That Fit MeAnswer a few questions, see your matches.Early WindowFound before the big boards.Direct ApplyStraight to the manager, past the ATS.By specialty
Your materials
CV AnalyzerWhat an ATS sees, and what to fix.Tailor CVBrought in line with one posting.Cover LetterWritten from your CV and the role.You and the process
Hey, I’m Wayjo. I find roles before the big boards.
Free to browse. An account unlocks the rest.
Jobs
All Tech JobsThe whole board, newest first.Roles That Fit MeAnswer a few questions, see your matches.Early WindowFound before the big boards.Direct ApplyStraight to the manager, past the ATS.$222,900 - 334,300/ year
The whole posting in a few lines. Sign up to read it here and on every role you open.
Sign Up to ReadYour work days are brighter here.
We’re obsessed with making hard work pay off, for our people, our customers, and the world around us. As a Fortune 500 company and a leading AI platform for managing people, money, and agents, we’re shaping the future of work so teams can reach their potential and focus on what matters most. The minute you join, you’ll feel it. Not just in the products we build, but in how we show up for each other. Our culture is rooted in integrity, empathy, and shared enthusiasm. We’re in this together, tackling big challenges with bold ideas and genuine care. We look for curious minds and courageous collaborators who bring sun-drenched optimism and drive. Whether you're building smarter solutions, supporting customers, or creating a space where everyone belongs, you’ll do meaningful work with Workmates who’ve got your back. In return, we’ll give you the trust to take risks, the tools to grow, the skills to develop and the support of a company invested in you for the long haul. So, if you want to inspire a brighter work day for everyone, including yourself, you’ve found a match in Workday, and we hope to be a match for you too.
The Data Platform and Observability Engineering (DPOE) team is building Workday's next-generation, multi-petabyte scale Observability Platform. We own the libraries, distributed services, and infrastructure that power ingestion, storage, and query across the observability stack — Iceberg, ClickHouse, Tempo, Grafana, S3, Kafka, and Elasticsearch — including LangSmith for LLM/agentic tracing and evaluation serving traces, metrics, and logs for every workload at Workday. Our roadmap directly shapes how the company detects, diagnoses, and eventually predicts operational issues at scale.
About the Role
To own the technical vision and architecture for distributed tracing as a first-class pillar of Workday's Observability Platform, built on ClickHouse and/or Grafana Tempo , backed by a big-data pipeline (Kafka, Spark/Flink, Iceberg, Clickhouse, Tempo,S3) running on AWS . This is a hands-on, high-autonomy role for an engineer who can design and build multi-petabyte, low-latency tracing infrastructure end-to-end — and who is equally excited to help define where Observability AI goes next: using traces, logs, and metrics as the substrate for automated root-cause analysis, anomaly detection, and AI-driven incident triage.
You'll set technical direction across multiple teams, mentor senior and staff engineers, and act as the primary architect and escalation point for the tracing subsystem — from ingestion and storage design through query performance and platform reliability.
Architect and build Workday's distributed tracing platform on ClickHouse/Tempo, designed for multi-petabyte scale ingestion and sub-second interactive query performance.
Own the big-data pipeline feeding tracing data — Kafka-based ingestion, Spark/Flink stream and batch processing, and Iceberg-on-S3 storage — including schema design, partitioning, compaction, and lifecycle management.
Drive performance and scaling across ingestion and query paths: storage format optimization (Parquet/Iceberg), compression strategy, partitioning/indexing, and query engine tuning under real production load.
Lead HA/DR design for tracing services — multi-region/multi-AZ resilience, failover, backup/restore, and recovery time/point objectives appropriate to a tier-1 platform.
Design security architecture for the platform, including authentication/authorization (authn/authz) for multi-tenant data access across ingestion and query layers.
Own operational excellence for distributed tracing: monitoring, logging, alerting, capacity planning, and participation in an on-call rotation for the platform.
Evaluate and introduce new technologies — open source and cloud-native — that materially improve the platform's scalability, cost efficiency, or capability.
Shape the future of Observability AI : Extend distributed tracing to LLM and agentic workflows using LangSmith and LangChain, enabling observability into multi-step agent execution, tool calls, and prompt/response chains.
Evangelize the platform : publish best practices, mentor engineers across DPOE and partner teams, and act as a technical thought leader for the modern observability/data stack internally.
Operate with high autonomy in a fast-moving, ambiguous environment — setting technical direction with minimal oversight while aligning with broader platform strategy.
Part of the week in the office
Workday
Office in United States
Also hiring in Auckland, New Zealand, Dublin, Ireland, Vancouver, Canada and 21 more places
5 of their 147 open roles are remote
Worth a look before you spend an evening tailoring a CV for it.
Still hiring?
14 checks1 red flag
How crowded?
7 checks2 red flags
How well does this role fit you?
Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.
Something wrong with this vacancy?