You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Be the first to open itNo views yet
Company hidden

Data Engineer

  • Office

Salary

Not stated

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

We're the fastest-growing startup transforming the IP industry.

Traction: 20-30% MoM revenue growth; selling to 700+ global IP teams (DLA Piper, tech boutiques, and global enterprises).

Proven Value: Users report 50-90% efficiency gains using our AI platform.

About the role

We’re hiring a data engineer to build the ingestion and search systems behind the company’s AI products.

Our sources include global patent literature, case law, technical standards and contributions, scientific databases, academic papers, and content from across the web. The data spans structured records, documents, images, audio and video. You’ll work across bulk ingestion and on-demand retrieval, making this information searchable and useful in our products

You’ll own systems from source acquisition through to serving queries. The work includes:

Large-scale ingestion. Build and operate high-throughput, resumable pipelines for large datasets, with efficient incremental updates, monitoring and recovery from failures.

Document processing and data quality. Extract useful content from complex documents and other formats. Handle malformed records and changing schemas, and validate outputs while preserving structure and metadata.

Search and serving. Build keyword, vector and structured search, and design schemas, indexes and partitioning for fast queries over tens to hundreds of millions of records.

Connecting information across sources. Link patents, scientific records and supporting documents, preserve dates and versions, and make results traceable to their original sources.

Performance engineering. Profile parsing, ingestion, database builds and queries throughout development, testing against representative datasets at realistic scale. Diagnose CPU, memory and storage I/O bottlenecks, and tune jobs and infrastructure for throughput, latency and cost.

You’ll work closely with our AI researchers and product engineers, with substantial freedom to choose the approach and build the systems yourself.

What you bring

Requirements

  • Strong Python and SQL, with experience designing and operating production databases.
  • Solid experience building and operating production data pipelines over large, messy datasets.
  • Expertise with running search systems over large document collections.
  • End-to-end ownership from raw data to user-facing functionality.
  • A good understanding of schema design, indexing and query optimisation.
  • A track record of diagnosing and fixing performance bottlenecks in live systems through profiling and measurement.
  • Experience with PostgreSQL/pgvector, OpenSearch (or Elasticsearch), Spark/Delta Lake, AWS, NoSQL databases, or Rust/C++ is useful.
  • The Founders
  • You'll partner with a founding team of AI PhDs and elite systems engineers:
  • Sanj (CRO): PhD in AI (Gatsby Unit, UCL), ex-Huawei R&D, former lead at Magic Carpet AI (acquired).
  • Chris (CEO): PhD in AI (UCL), published researcher, ex-Dyson and Alan Turing Institute.
  • Angus (CTO): MEng Computer Science, ex-Qualcomm and Coremont (Brevan Howard).

Benefits

  • Competitive Salary + Significant Equity: We want you to have true ownership in the success of the company.
  • Founding Impact: You'll have a direct hand in how we build out the data infrastructure the rest of the product depends on.
  • Support: Full visa sponsorship and private medical insurance.
  • The Environment: Free meals and a seat at the table with an incredibly smart, ambitious team.
  •  
  • Experience: Any (new grads ok)
  • Visa: US citizen/visa only

Where you’d work

From the office

No visa sponsorship

About the company

Company hidden

Office in London, United Kingdom

Your chances

Still hiring, moderately crowded, and a person reads your message.

  • 16 checks run
  • 4 good signs
  • 1 red flag

Still hiring?

11 checks

Actively hiring

In its favour2

  • Still on the company's own careers site, checked 3 h agoModerate evidence
  • A hiring contact is attached to itSlight evidence

How crowded?

5 checks

Moderate

In its favour2

  • You can message the hiring contact and skip the queueModerate evidence
  • In the office in London: only people nearby can take itSlight evidence

Against it1

  • Open for 2 weeks: applications have had time to pile upModerate evidence
8 roles like this are in their Early WindowFound before the big job boards, while the crowd hasn’t arrived

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details11 facts · Role, Location, Compensation, Employment
Tech stack
  • Python
  • SQL
  • PostgreSQL
  • AWS
  • Spark
  • Vector databases
Type
Full-time
Equity
Equity offered
Specialty
Data Engineering
Region
United Kingdom
Pay period
Annual
Show 5 more factsShow less

Role

Category
Data & Analytics
Specialty
Data Engineering
Tech stack
  • Python
  • SQL
  • PostgreSQL
  • AWS
  • Spark
  • Vector databases

Location

Work model
Office
Region
United Kingdom
Office
  • London, United Kingdom
Visa sponsorship
Not sponsored

Compensation

Salary
Salary by agreement
Pay period
Annual
Equity
Equity offered

Employment

Type
Full-time

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Data Engineer

    $345,000 - 385,000 / year

    • Hybrid · Seattle
    • Mid-Level
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Lead Data Engineer

    $110,000 - 180,000 / year

    • Office · US
    • Lead & Manager
    Direct apply
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior Data Engineer

    $200,000 - 250,000 / year

    • Office · San Francisco
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Technical Lead, Data Platform Engineer

    Salary by agreement

    • Hybrid · London
    • Lead & Manager

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason