Observed Signal · Jul 22, 2026 · Technical Release · Source: DEV Community · Impact: 2/5 · Sentiment: Positive
One API for Company Job Postings
The article demonstrates that most tech companies use a small set of Applicant Tracking Systems (Greenhouse, Lever, Ashby, Workable, SmartRecruiters, Recruitee) which expose public JSON endpoints. Because of that, job-posting collection is primarily a normalization problem rather than classic HTML scraping. The author shows an Apify actor (fetchbase job-postings-scraper) that auto-detects a company's ATS and returns a single normalized record per job, with honest nulls where sources omit fields like remote or salary. The piece highlights that some ATS boards (notably Ashby) publish fully structured compensation ranges (e.g., OpenAI’s Ashby board: 727 roles all include pay ranges, 460 flagged remote).
Provides a practical, low-friction way to collect normalized job and structured compensation data from public ATS APIs — useful for labor-market signals and HR/people-data projects but not transformational for the AdTech/MarTech industry.
Track Stripe Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- Common ATS platforms mentioned: Greenhouse, Lever, Ashby, Workable, SmartRecruiters, Recruitee.
- Apify's Job Postings actor (fetchbase job-postings-scraper) auto-detects ATS and returns one normalized row per job.
- At time of writing the article reports 518 open roles at Stripe, 280 at Palantir, and 727 at OpenAI.
- OpenAI's Ashby board reportedly published pay ranges for all 727 roles and flagged 460 of them as remote.
- Apify bills per job returned and the actor cannot invent fields omitted by the source ATS (e.g., remote, salary may be null).
Connected Companies & Entities
4 Entities mapped“At the time of writing that's 518 open roles at Stripe, 280 at Palantir, and 727 at OpenAI — three calls, three completely different JSON sh...”
“At the time of writing that's 518 open roles at Stripe, 280 at Palantir, and 727 at OpenAI — three calls, three completely different JSON sh...”
“At the time of writing that's 518 open roles at Stripe, 280 at Palantir, and 727 at OpenAI — three calls, three completely different JSON sh...”
“Example API request includes companies: 'companies': ['stripe', 'gitlab', 'https://jobs.ashbyhq.com/openai']...”
Ontology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
Aggregators block scrapers; ATS job APIs are public
An analysis published on 2026-08-01 found that major job aggregators (Indeed, Glassdoor, ZipRecruiter, Upwork) often return 403 responses with empty bodies (blocked by CDNs/WAFs), while the original sources—company career boards served by Applicant Tracking Systems (ATS)—expose public JSON job-board APIs that answer anonymous requests with no API key. The author measured responses from Greenhouse, Ashby, Lever, Rippling, Workable, Recruitee and SmartRecruiters and found these endpoints return clear semantics (200 with jobs, 200 empty list for no openings, 404 for unknown board) in most cases, though SmartRecruiters behaves differently. Practical notes include the need to discover board tokens, differences in payload size when requesting full content (Greenhouse descriptions increase payload ~18x), and that some ATS endpoints publish compensation bands behind flags.
AI Recruiter Using RAG, Memory and Web Search
Recruit Intelligence Agent is an open developer project — an AI-powered recruitment assistant built with FastAPI and Backboard. It provides resume parsing into a JSON Resume schema, automated candidate screening and scoring, job-description generation with market research, memory-enabled multi-step (agentic) reasoning pipelines, and candidate validation via live web search. The repository and installation instructions are published on GitHub; the app exposes REST endpoints (upload, parse, evaluate, comprehensive_evaluate, websearch, validate, jd generation, summarization, QA) and supports stateful (thread-based) and stateless modes. The project was submitted to an Earth Day Hackathon 2026 and demonstrates integration of RAG, persistent memory, and web-validation in an end-to-end hiring API.
Developer Releases AI Web Data Extractor API
A developer published an AI Web Data Extractor API that combines fast HTTP scraping (Axios + Cheerio) with a Puppeteer browser fallback to extract structured data from arbitrary URLs. Implemented in Node.js, the extractor can return product data (title, price, image), emails, and article metadata, and uses a heuristic to auto-fallback to browser rendering when static scraping yields weak results. The API is available via RapidAPI and the author provides code snippets (fetchStatic, fetchBrowser, extractProduct) plus an example POST request/JSON response. The post lists real use cases (SaaS, price tracking, lead generation), implementation challenges (anti-bot measures, messy price formats), and planned additions such as proxy rotation, CAPTCHA bypass, and LLM-based parsing and page classification.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
