- Company: Get A Job.ai
- Location: London
- Salary: Pay not listed
Website Get A Job.ai
Represented by Get A Job.ai
Responsibilities
We are representing a confidential AI infrastructure organization seeking a Staff or Senior Software Engineer to lead the development of web-scale crawling and content acquisition systems for a novel search platform designed specifically for AI agents.
In this role, you will build distributed systems that discover, fetch, and continuously refresh content from the open web and large-scale data sources. Your work will enable AI systems to access fresh, verified real-world information through high-performance search APIs.
Your core responsibilities include:
- Design, implement, and operate web-scale crawling systems capable of processing billions of URLs
- Build ingestion workflows for diverse data sources including crawlers, structured feeds, and partner integrations
- Develop crawl scheduling, prioritization, recrawl policies, and freshness strategies to balance coverage and resource efficiency
- Build systems for URL discovery, deduplication, content extraction, and crawl orchestration
- Ensure reliable operation of high-throughput crawling infrastructure at internet scale
- Define and monitor observability metrics for crawl coverage, freshness, throughput, and content quality
- Optimize resource usage, bandwidth consumption, and infrastructure costs
- Collaborate with indexing and machine learning teams to ensure content quality meets downstream requirements
- Enable safe experimentation with crawling strategies and acquisition policies
What We're Looking For
Required qualifications:
- 5+ years of experience building backend or distributed systems
- Strong expertise in Go or C++
- Proven experience with large-scale distributed systems handling 10,000+ requests per second, billions of URLs, or high-throughput data pipelines
- Deep understanding of web protocols including HTTP, DNS, and TLS
- Hands-on experience with crawling, scraping, and content extraction at scale
- Track record operating production systems and debugging failures in distributed environments
- Strong grasp of scalability, fault tolerance, and resource management principles
Preferred experience:
- Direct web crawling experience
- Building streaming data pipelines and event-driven architectures
- Working with messaging platforms such as Kafka, Pulsar, NATS, or RabbitMQ
- Designing distributed schedulers, queues, and asynchronous processing systems
- Experience with Spark, Flink, Beam, or MapReduce frameworks
- Background in ad tech, social networks, search engines, or other large-scale content platforms
This position is based in London. A coding interview will be part of the selection process. Applicants must be authorized to work in the United Kingdom.
How We Work With You
As a recruiting partner for this confidential client, Get A Job.ai manages the full candidate experience. When you apply through our platform at getajob.ai, one of our specialist recruiters will review your background and conduct an initial screening call. If there's a strong match, we'll submit your profile to our client for consideration and coordinate all interview stages.
Please apply exclusively through Get A Job.ai. Do not attempt to contact the client directly, as this may disqualify your application.
Pay
Compensation details will be discussed during the screening process with our recruiting team.
Equal Opportunity: Get A Job.ai and our clients are committed to fostering inclusive, diverse workplaces. We provide equal employment opportunities without regard to race, color, religion, sex, national origin, age, disability, veteran status, sexual orientation, gender identity, or any other protected characteristic.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Explore Get A Job.ai online
Working in London, UK
Weather right now in London, UK: checking… · Local time: · Air quality: · Daylight: · UV index: · Wind: · Pollen:
London is the capital and largest city of England and the United Kingdom, with a population of 9.1 million people in 2024. Its wider metropolitan area is the largest in Western Europe, with a population of 15.4 million. London stands on the River Thames in southeast England, at the head of a 50-mile (80 km) tidal estuary down to the North Sea, and has been a major settlement for nearly 2,000 years. Its ancient core and financial centre, the City of London, was founded by the Romans as Londinium and has retained its medieval boundaries. The City of Westminster, to the west of the City of London
England is a country that is part of the United Kingdom. It is located on the island of Great Britain, of which it covers about 62%, and more than 100 smaller adjacent islands. England shares a land border with Scotland to the north and another land border with Wales to the west, and is surrounded by the North Sea to the east, the English Channel to the south, the Celtic Sea to the south-west, and
Nearby green space: 12 parks within 1.5km — closest is Whitehall Garden (364m). via OpenStreetMap
Nearest public transit: Charing Cross (station, 41m). via OpenStreetMap
- Elevation 18m (59 ft)
Source: Wikipedia (state)
Job details above are provided by the employer/source. The sections on this page are compiled from public data sources with AI assistance.
Accommodations: if you need a workplace accommodation to apply for or perform this job, see ADA.gov or EEOC.gov for guidance on your rights and how to request one.
Listing facts
- Role Staff / Senior Software Engineer (Agentic Search) – Crawler
- Employer Get A Job.ai
- Location London
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 15, 2026
- Apply by October 15, 2026
- Overview Full job description on this page (488 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
