Our job scraper was nearly 5x slower and costlier because of a proxy it never needed
We built a tool that takes company names and returns their open jobs: type Databricks, OpenAI, Lyra Health, get their postings as JSON. Behind it is an index of more than 500 companies whose job boards on Greenhouse, Lever and Ashby we had proven belong to them (how, and why guessing gets it wrong half the time, is in an earlier post). The first runs on the real platform found three problems that no unit test had. All three returned HTTP 200, and all three had been sitting in tools that were already public. Input: three companies, 60 jobs.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in