Staff / Senior Software Engineer (Agentic Search) - Crawler

🗓️ Posted 2026-09-01 Amsterdam, Netherlands; London, United Kingdom Hybrid ICT

Company shared salary

NA

Market rate

6,000 GBP–9,500 GBP/mo (72,000 GBP–114,000 GBP/yr)

Based on similar roles (title + domain + location).

About the Company

Nebius is leading a new era in cloud infrastructure for the global AI economy. They are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers, they own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, they have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Their team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. In a rapidly evolving world, trust in AI depends on AI agents being grounded in fresh, verified real-world data. They are building an agent-native search platform designed specifically for AI systems rather than human users. Their product provides programmatic, low-latency, and observable search APIs that AI agents use to retrieve, filter, and reason over real-world information at scale.

About the Role

We are looking for a Senior Software Engineer to work on the content acquisition and crawling infrastructure of a novel search engine tailored for agentic AI consumption. In this role, you will focus on building systems that discover, fetch, and continuously refresh content from the open web and other large-scale data sources. You will design distributed crawling, scheduling, and ingestion infrastructure capable of operating at internet scale while balancing coverage, freshness, resource efficiency, and reliability. You will work on systems that process billions of URLs, manage high-throughput data flows, and ensure that high-quality content is consistently available to downstream indexing and retrieval systems.

Responsibilities

  • Design, implement, and operate web-scale crawling systems for acquiring content from the internet.
  • Build ingestion workflows for internal and external data sources, including crawlers, structured feeds, and partner integrations.
  • Develop crawl scheduling, prioritisation, recrawl policies, and freshness strategies.
  • Build systems for URL discovery, deduplication, content extraction, and crawl orchestration.
  • Ensure reliable operation of crawling infrastructure under high-throughput conditions.
  • Define observability and quality metrics for crawl coverage, freshness.