Role descriptionGelesen
About the Job We build and operate l arge-scale web crawling and extraction systems that turn unstructured, dynamic websites into clean, structured, and correct data. You'll own crawlers end-to-end: discovery, resilient fetching against real anti-bot defenses, and turning HTML/PDF into trustworthy structured output. What you will do • Build and scale crawlers that handle dynamic, large-scale sites - and keep them running as those sites change and fight back. • Stay ahead of anti-bot measures: TLS/browser fingerprinting, proxies, session and rate-limit strategy, change detection. • Turn raw HTM…
Read the full description & apply at the source →