Job posting feeds
Daily or hourly collection of new postings for a set of titles, locations or companies, deduplicated across boards and employer sites, with first-seen and last-seen dates per job so expiry is measurable.
Industries
Job postings, company career pages, salary pages and employer reviews, deduplicated into one feed with posting and expiry dates.
Job data is the vertical we have scraped the longest. Postings are duplicated across boards, expire silently, hide salary in free text and sit behind some of the strongest anti-bot setups on the web. We collect from boards, aggregators and employer career sites, parse titles, locations, salary ranges and skills into fields, deduplicate the same job across sources and record when each posting appeared and disappeared. Candidate profiles and anything behind a login are not part of the work.
Sources
Public pages and documents we have scraped in this vertical. Named sites are examples, not an exhaustive list.
Schema
A common starting schema. You decide the final columns and names; we keep them stable across runs.
job_idtitlecompanylocationremotesalary_minsalary_maxcurrencyemployment_typeposted_atsourcescraped_atOutput
job_id | title | company | location | remote | salary_min | salary_max | currency | employment_type | posted_at | source | scraped_at |
|---|---|---|---|---|---|---|---|---|---|---|---|
| JB-00912377 | Senior backend engineer | Example Corp | Berlin, DE | hybrid | 70000 | 90000 | EUR | full-time | 2026-10-01 | Employer site | 2026-10-07 |
Values are illustrative placeholders to show shape and types, not records from any client or source.
Use cases
Daily or hourly collection of new postings for a set of titles, locations or companies, deduplicated across boards and employer sites, with first-seen and last-seen dates per job so expiry is measurable.
Parse salary ranges, seniority and required skills from posting text into structured fields across thousands of postings, so compensation and demand can be studied by role, city and employer over time.
Track open roles per company over time from career sites and boards, giving a public view of where companies are hiring, which teams are growing and which postings stay open the longest.
Questions
Public job and company pages, within the limits those sites place on access and without logging in. Member profiles, messaging and anything behind a login are out of scope.
On employer, normalised title, location and posting text similarity, with the source identifiers kept on every record. You see one job with a list of where it was found.
Yes. Ranges, currencies and pay periods are parsed into numeric fields with a flag for estimated versus stated, and the original text is kept for review.
Most job clients run daily. Each run lands in your storage with a diff of new, changed and removed postings, priced per month after the first sample.
Send the sources and the fields you need. An engineer looks at them the same day and replies with a plan and a price.
Hi! Send me the site you need data from and I'll get an engineer to look at it.
Chat on WhatsApp Or get a quote