Still hiring?
11 checksActively hiring
In its favour2
- Still on the company's own careers site, checked 3 h agoModerate evidence
- A hiring contact is attached to itSlight evidence
Browse
All Tech JobsThe whole board, newest first.Roles That Fit MeAnswer a few questions, see your matches.Early WindowFound before the big boards.Direct ApplyStraight to the manager, past the ATS.By specialty
Your materials
CV AnalyzerWhat an ATS sees, and what to fix.Tailor CVBrought in line with one posting.Cover LetterWritten from your CV and the role.You and the process
Hey, I’m Wayjo. I find roles before the big boards.
Free to browse. An account unlocks the rest.
Jobs
All Tech JobsThe whole board, newest first.Roles That Fit MeAnswer a few questions, see your matches.Early WindowFound before the big boards.Direct ApplyStraight to the manager, past the ATS.Not stated
The whole posting in a few lines. Sign up to read it here and on every role you open.
Sign Up to ReadWe're the fastest-growing startup transforming the IP industry.
Traction: 20-30% MoM revenue growth; selling to 700+ global IP teams (DLA Piper, tech boutiques, and global enterprises).
Proven Value: Users report 50-90% efficiency gains using our AI platform.
About the role
We’re hiring a data engineer to build the ingestion and search systems behind the company’s AI products.
Our sources include global patent literature, case law, technical standards and contributions, scientific databases, academic papers, and content from across the web. The data spans structured records, documents, images, audio and video. You’ll work across bulk ingestion and on-demand retrieval, making this information searchable and useful in our products
You’ll own systems from source acquisition through to serving queries. The work includes:
Large-scale ingestion. Build and operate high-throughput, resumable pipelines for large datasets, with efficient incremental updates, monitoring and recovery from failures.
Document processing and data quality. Extract useful content from complex documents and other formats. Handle malformed records and changing schemas, and validate outputs while preserving structure and metadata.
Search and serving. Build keyword, vector and structured search, and design schemas, indexes and partitioning for fast queries over tens to hundreds of millions of records.
Connecting information across sources. Link patents, scientific records and supporting documents, preserve dates and versions, and make results traceable to their original sources.
Performance engineering. Profile parsing, ingestion, database builds and queries throughout development, testing against representative datasets at realistic scale. Diagnose CPU, memory and storage I/O bottlenecks, and tune jobs and infrastructure for throughput, latency and cost.
You’ll work closely with our AI researchers and product engineers, with substantial freedom to choose the approach and build the systems yourself.
What you bring
From the office
No visa sponsorship
Company hidden
Office in London, United Kingdom
Still hiring, moderately crowded, and a person reads your message.
Still hiring?
11 checksActively hiring
In its favour2
How crowded?
5 checksModerate
In its favour2
Against it1
How well does this role fit you?
Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.
Something wrong with this vacancy?