You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Be the first to open itNo views yet
ClickHouse

Cloud Software Engineer - Observability Platform

  • Remote
  • 3-6 years

Salary

Not stated

Similar roles pay $145K - 190K a year · our estimate

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

ClickHouse is looking for an experienced engineer to join our Observability organization. We are hiring across two closely aligned teams: Observability Platform and Internal Observability .

Together, these teams build and operate the systems behind ClickHouse’s internal observability and ClickStack Cloud. Our platforms process trillions of events per day, sustaining throughput in the hundreds of millions of events per second. The experience we gain operating observability at ClickHouse scale feeds directly back into the platform and product we provide to customers.

The Observability Platform team builds shared systems for telemetry ingestion, durable buffering, processing, storage, autoscaling, and service provisioning. The Internal Observability team operates ClickHouse’s company-wide observability platform and works with engineering teams to improve reliability, debugging, and operational efficiency.

This is a software engineering role at the intersection of distributed systems, cloud infrastructure, and production operations. You will build systems, operate what you build, respond to incidents, and turn recurring operational problems into durable software and automation.

Your background may be in backend engineering, infrastructure, SRE, or systems engineering. What matters most is your ability to solve ambiguous production problems, make sound engineering tradeoffs, and take ownership from design through operation.

Responsibilities

  • Design, build, and operate distributed systems that ingest, process, and store telemetry at very high scale.
  • Own the reliability, performance, capacity, and cost-efficiency of telemetry pipelines and storage systems.
  • Participate in the on-call rotation, help resolve production incidents, and drive root-cause fixes through to completion.
  • Build software and automation that eliminate repetitive operational work and make the platform easier to operate.
  • Identify architectural bottlenecks and help shape the roadmap for the next stage of scale.
  • Work closely with product, infrastructure, and service teams across ClickHouse.
  • Contribute to architecture and design reviews and help raise engineering quality across the team.
  • How you work
  • You take ownership and proactively improve the systems around you.
  • You are comfortable debugging unfamiliar distributed systems in production.
  • You make pragmatic tradeoffs between reliability, performance, delivery speed, and cost.
  • You communicate clearly in a remote, async-friendly environment.
  • You prefer incremental delivery: ship a focused solution, validate it in production, and improve it based on evidence.
  • You address the underlying cause of operational problems rather than repeatedly treating the symptoms.
  • Technical experience
  • 5+ years of experience building and operating production systems at scale.
  • Strong proficiency in Go.
  • Experience building and operating services on Kubernetes.
  • Experience with infrastructure-as-code and GitOps tooling such as Terraform, Helm, and Argo CD.
  • Production experience with at least one major cloud provider: AWS, GCP, or Azure.
  • Hands-on experience with telemetry systems such as OpenTelemetry, Prometheus, Grafana, or comparable technologies.
  • Bonus points
  • Experience with ClickHouse.
  • Experience with high-throughput ingestion, streaming, queueing, or storage systems.
  • Experience building multi-tenant cloud services.
  • Experience optimizing infrastructure for both performance and cost.
  • Experience with TypeScript.
  • #LI-Remote

Benefits

  • Flexible work environment - ClickHouse is a globally distributed company and remote-friendly. We currently operate in over 25 countries.
  • Healthcare - Employer contributions towards your healthcare.
  • Equity in the company - Every new team member who joins our company receives stock options.
  • Time off - Flexible time off in the US, generous entitlement in other countries.
  • A USD$500 Home office setup if you’re a remote employee.
  • Global Gatherings – We believe in the power of in-person connection and offer opportunities to engage with colleagues at company-wide offsites.
  • Culture - We All Shape It
  • As part of a rapidly scaling start-up, you will be instrumental in shaping our culture.
  • Are you interested in finding out more about our culture? Learn more about our values here . Check out our blog posts or follow us on LinkedIn to find out more about what’s happening at ClickHouse.
  • Equal Opportunity & Privacy
  • ClickHouse provides equal employment opportunities to all employees and applicants and prohibits discrimination and harassment of any type based on factors such as race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
  • Please see here for our Privacy Statement.

Where you’d work

Fully remote

You can work from

  • United States

About the company

ClickHouse

  • Industry: SaaS

Also hiring in Portugal, United Kingdom, Canada and 9 more places

41 of their 47 open roles are remote

Your chances

Worth a look before you spend an evening tailoring a CV for it.

  • 19 checks run
  • 3 red flags

Still hiring?

13 checks

1 red flag

How crowded?

6 checks

2 red flags

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details11 facts · Role, Location, Compensation, Employment, Company
Tech stack
  • Go
  • AWS
  • Kubernetes
  • Terraform
  • TypeScript
  • ClickHouse
  • Azure
  • Argo CD
  • Helm
  • Prometheus
  • Grafana
Type
Full-time
Equity
Equity offered
Industry
SaaS
Region
United States
Pay period
Annual
Show 5 more factsShow less

Role

Category
DevOps & Infrastructure
Experience
5+ years
Tech stack
  • Go
  • AWS
  • Kubernetes
  • Terraform
  • TypeScript
  • ClickHouse
  • Azure
  • Argo CD
  • Helm
  • Prometheus
  • Grafana

Location

Work model
Remote
Region
United States
Remote from
  • United States

Compensation

Salary
Salary by agreement
Pay period
Annual
Equity
Equity offered

Employment

Type
Full-time

Company

Industry
SaaS

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Lead IT Engineer

    Salary by agreement

    • Remote
    • Lead & Manager
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Security Engineer, New Grad

    Salary by agreement

    • Office · Dublin
    • Junior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Site Reliability Engineer

    Salary by agreement

    • Remote
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior Security Operations Analyst

    Salary by agreement

    • Hybrid · Wellington, Auckland
    • Senior

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason