ActorStack.dev

The startup data this Actor will not collect

Founders, employees, funding rounds and investors are all on Wellfound and none of them are in the output. What the line is, and why it sits where it does.

By Oswaldo Carabano6 min read

Short answer

Wellfound company profile pages carry funding rounds, investors, perks and named founders and employees. This Actor reads none of them. The immediate reason is that Wellfound puts `/company/*` behind a Cloudflare challenge; the reason that would apply anyway is that the page identifies individuals, and a job-market dataset does not need to name people to be useful. What ships instead is the company as it appears attached to its own postings — pitch, size band, badges, industries, website and geocoded headquarters — plus job descriptions delivered exactly as the company wrote them, with nothing extracted or derived from them.

Key points

  • Company profile pages are behind a Cloudflare challenge, so reading them would mean defeating a control the site put there deliberately.
  • Those pages also name founders and employees, which turns a market dataset into a file about individuals.
  • Job descriptions are returned exactly as written and nothing is extracted or derived from them — no names, no contact details, no inferred seniority.
  • The company information that does ship comes from the posting surface, which is published for candidates to read.
  • wellfound.com/robots.txt separately disallows `/u/`, the user profile path, which points the same way.
On this page5 sections

Wellfound holds a good deal more than job postings. This is an account of what was left out and why, written because a limits section that only lists accidents is not a limits section.

What is on a company page

Funding rounds, investors, perks, team size and — the part that matters here — named founders and employees with their roles. It is a rich page, and it is the page that gets asked about most.

Two reasons, pointing the same way

The first is that Wellfound puts /company/* behind a Cloudflare challenge. That is a control the site placed there deliberately, and treating it as an obstacle to route around is a different activity from reading published listings.

The second would apply even if the page were wide open: it identifies individuals. A dataset about which startups are hiring for what, at what salary, does not become more useful by naming the people who work there — it becomes a different kind of dataset, with a different set of obligations attached. Article 6 GDPR does not exempt data because it was published.

Descriptions are passed through, not mined

Job descriptions are delivered exactly as the company wrote them, in markdown and HTML. Nothing is extracted from them and nothing is derived: no hiring-manager names, no contact details pulled out of the text, no inferred seniority. A description that happens to contain a name contains it because the company published it there.

What ships instead

The company as it appears attached to its own postings: pitch, size band, hiring badges, industries, website, logo and geocoded headquarters, plus every ATS seen across its jobs. That is assembled from the listing surface, which exists so that candidates can find these companies.

What robots.txt says about profiles

wellfound.com/robots.txt, checked on 9 September 2026, disallows /u/ — the user profile path — along with /search and the account surfaces. A site's own machine-readable declaration is the cheapest signal available about what it considers crawler-facing, and here it points the same direction as everything else — reading robots.txt as a specification.

Frequently asked questions

Can I get founder or employee names?
Not from this Actor. Company profile pages, which is where those appear, are behind a Cloudflare challenge and are not read at all. The output describes companies and their vacancies rather than the people at them.
Are job descriptions analysed?
No. They are delivered exactly as the company wrote them, in markdown and HTML. Nothing is extracted or derived from the text — no names, no contact details, no inferred seniority.
Is funding data available?
No. Funding rounds, investors and perks all live on the company profile page, which this Actor does not read.

Sources

Every URL below was requested and returned a page on the date shown.

  1. Site declarationchecked 9 Sept 2026
    wellfound.com/robots.txtWellfound
  2. Operator claimchecked 9 Sept 2026
    Wellfound Terms of ServiceWellfound
  3. Law or regulatorchecked 9 Sept 2026
    Art. 6 GDPR — Lawfulness of processingRegulation (EU) 2016/679
  4. Operator claimchecked 9 Sept 2026
    Wellfound Jobs Scraper — Actor README and input schemaActorStack / Apify Store
A laptop screen showing a plain text-mode terminal with a command prompt.
WellfoundExplainer

The Wellfound API question

There is no documented public endpoint for startup job listings. What robots.txt permits, what Cloudflare sits in front of, and which surface is actually readable.

6 min
A white measuring tape curving across a dark background, showing the numbers 15 to 45.
WellfoundExplainer

The column not shipped

Wellfound exposes a maximum years-of-experience value. It was filled on 2 of 2,229 jobs. Shipping it would have added a field that looks like data and is not.

5 min
Developers working side by side at desktop computers in an open-plan tech office.
WellfoundGuide

Scrape Wellfound jobs

A walkthrough of extracting startup job listings: how Wellfound's URL-path filtering constrains what you can ask for, which caps actually bound a run, and what arrives in each row.

9 min