Website Job Extractor

The NanoScrape Website Job Extractor Scraper is an Apify actor that extracts enrichment data from Website Job Extractor. It returns 25 fields per result including title, location, description, company_name, salary_range, priced at $10/1k company processed, with no API key or monthly subscription required.

Website Job Extractor (Career Pages, Global): AI-powered extraction from any company career page. Auto-detects ATS, handles pagination, and validates output. Freshest hiring-intent signal for B2B targeting.

Open on Apify →
Pricing
$10.00 / 1,000 company processed
Runtime
Cloud (Apify)
Proxy
Datacenter

What you can scrape — 25 fields per result

FieldTypeExampleGroupFill %
company_idstring2a59bd88-6ee5-4082-9da0-26ed3e5e9a30Core100%
company_namestringNASH & CO SOLICITORS LLPCore100%
titlestring | nullCommercial Property SolicitorCore31%
source_urlstringhttps://www.careers.nash.co.uk/Core37%
job_urlstring | nullhttps://www.careers.nash.co.uk/vacancy-commercial-property-solicitorCore12%
job_idstring | nullvacancy-commercial-property-solicitorCore8%
application_urlstring | nullhttps://www.careers.nash.co.uk/Core31%
ats_urlstring | null(populated when data available)Core5%
career_page_urlstring | nullhttp://www.nash.co.ukCore31%
locationstring | nullPlymouthLocation21%
descriptionstring | nullExperienced solicitor needed for the specialist Commercial Property team at Nash & Co S...Description25%
salary_rangestring | null(populated when data available)Salary10%
employment_typestring | nullFull-timeEmployment18%
experience_levelstring | null(populated when data available)Requirements16%
requirementsarray(populated when data available)Requirements15%
benefitsarray(populated when data available)Requirements11%
departmentstring | nullCommercial PropertyMetadata22%
posted_atstring | null(populated when data available)Metadata6%
workplace_typestring | null(populated when data available)Other17%
extracted_atstring2026-08-19T07:07:06.253ZOther100%
confidencenumber | null0.6500000000000001Other28%
relevancestring | null(populated when data available)Other
ats_systemstring | null(populated when data available)Other5%
js_rendering_suspectedboolean(populated when data available)Other6%
js_indicatorsarray(populated when data available)Other6%

Output example — real dataset item (PII scrubbed)

{
  "company_id": "2a59bd88-6ee5-4082-9da0-26ed3e5e9a30",
  "company_name": "NASH & CO SOLICITORS LLP",
  "title": "Commercial Property Solicitor",
  "description": "Experienced solicitor needed for the specialist Commercial Property team at Nash & Co Solicitors in Plymouth, handling commercial property matters.",
  "location": "Plymouth",
  "employment_type": "Full-time",
  "department": "Commercial Property",
  "source_url": "https://www.careers.nash.co.uk/",
  "job_url": "https://www.careers.nash.co.uk/vacancy-commercial-property-solicitor",
  "job_id": "vacancy-commercial-property-solicitor",
  "application_url": "https://www.careers.nash.co.uk/",
  "extracted_at": "2026-08-19T07:07:06.253Z",
  "confidence": 0.6500000000000001,
  "ats_system": null,
  "ats_url": null,
  "career_page_url": "http://www.nash.co.uk"
}

Cost math

Priced at $0.0100 per company processed. 10 company processeds ≈ $0.10, 100 ≈ $1.00, 1,000 ≈ $10.00. The first runs land inside Apify's $5/mo free-tier credit — pay only for what you extract, no monthly subscription.

Who should use this

B2B sales and lead generation teams use this actor to detect hiring intent at the exact moment a company posts a new role. A company posting 5 new sales engineer roles is a live buying signal for CRM, HR tech, and benefits vendors, and this actor surfaces it directly from the company's own website before it appears on public job boards.

Recruitment analytics platforms building competitive intelligence feeds combine this actor with the Google Maps Scraper to discover companies by location, then automatically extract their open jobs for sector-level hiring trend dashboards.

Niche job board operators covering specific industries use the skipAtsExtraction: true mode to route Personio, Greenhouse, and Lever portals into the Career Site Jobs Scraper for structured API-level extraction, while the LLM handles the remaining custom career pages, all in one pipeline.

Integrations

Run this actor directly on Apify (no code)

Click Open on Apify above to run website-job-extractor in your browser - no code, no install. In the Apify console you get:

Get your Apify API token

To run this actor from your own code you need an Apify API token. Get one in about a minute:

  1. Sign up for a free Apify account (Google, GitHub, or email).
  2. Go to Settings → Integrations → API tokens.
  3. Click Create a new API token. Copy it and keep it secret.

Free tier: Apify credits your account with $5 of platform usage every month, no credit card required. Enough to test any actor meaningfully - at $0.001 per result on typical scrapers, that is roughly 5,000 results for free every month.

Call from code

Run this actor from any language via the Apify REST API. Replace YOUR_TOKEN with your API token and adapt the input JSON to your needs.

curl -X POST "https://api.apify.com/v2/acts/santamaria-automations~website-job-extractor/run-sync-get-dataset-items?token=YOUR_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{}'
// npm install apify-client
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: 'YOUR_TOKEN' });
const run = await client.actor('santamaria-automations/website-job-extractor').call({});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);
# pip install apify-client
from apify_client import ApifyClient

client = ApifyClient('YOUR_TOKEN')
run = client.actor('santamaria-automations/website-job-extractor').call(run_input={})
items = list(client.dataset(run['defaultDatasetId']).iterate_items())
print(items)
// dotnet add package Apify.Client
using Apify.Client;

var client = new ApifyClient("YOUR_TOKEN");
var run = await client.Actor("santamaria-automations/website-job-extractor").CallAsync(new { });
var items = await client.Dataset(run.DefaultDatasetId).ListItemsAsync();
// Maven: com.apify:apify-client
import com.apify.client.ApifyClient;

ApifyClient client = new ApifyClient("YOUR_TOKEN");
ActorRun run = client.actor("santamaria-automations/website-job-extractor").call(Map.of());
List<Map<String,Object>> items = client.dataset(run.getDefaultDatasetId()).listItems();

Use with AI agents (MCP)

This actor is available on the Apify MCP server, so you can drive it from any MCP-compatible AI client - Claude Desktop, Claude.ai, Cursor, VS Code, LangChain, LlamaIndex, or a custom agent - without writing any code.

https://mcp.apify.com?tools=santamaria-automations/website-job-extractor

Example prompt once connected:

"Use <code>website-job-extractor</code> to run a scrape on my target list and give me the results as a table."

Clients that support dynamic tool discovery (Claude.ai, VS Code) receive the full input schema automatically via add-actor.

Use with no-code platforms

Trigger this actor from your favorite automation tool. Every platform below can call the Apify API in a few clicks - no code required.

  1. Add the n8n Apify node (installed by default on n8n Cloud; on self-hosted install @apify/n8n-nodes-apify).
  2. Set Actor to santamaria-automations/website-job-extractor.
  3. Choose operation Run Actor and get dataset, paste your input JSON, and connect a downstream node (Google Sheets, Postgres, webhook, ...).
  1. Create a new Zap with any trigger.
  2. Add a Webhooks by Zapier POST action to https://api.apify.com/v2/acts/santamaria-automations~website-job-extractor/run-sync-get-dataset-items?token=YOUR_TOKEN.
  3. Body: your input JSON. Content-Type: application/json.
  1. Create a new scenario.
  2. Add an HTTP module: URL https://api.apify.com/v2/acts/santamaria-automations~website-job-extractor/run-sync-get-dataset-items?token=YOUR_TOKEN, method POST, body raw JSON.
  3. Parse the response with a JSON module to iterate items downstream.
  1. Create a new workflow.
  2. Add an HTTP Request step: POST to https://api.apify.com/v2/acts/santamaria-automations~website-job-extractor/run-sync-get-dataset-items?token=YOUR_TOKEN.
  3. Access steps.http.$return_value in subsequent steps.
  1. Install the API Connector plugin.
  2. Add a new API: POST https://api.apify.com/v2/acts/santamaria-automations~website-job-extractor/run-sync-get-dataset-items?token=YOUR_TOKEN, body type JSON.
  3. Initialize the call with your input, then reference the response array in workflow actions.

FAQ

How much does the website-job-extractor cost?

$10.00 / 1,000 company processed, billed per result on Apify. Apify's $5/month free tier covers the first runs. No monthly subscription — you only pay for what you extract.

What fields does it return?

25 fields per result. Full field catalog is listed in the 'What you can scrape' table above — includes core identifiers, descriptive text, dates, and any category-specific data.

Is scraping this site legal?

Scraping publicly available data is generally permitted in most jurisdictions, but GDPR applies to any personal data you collect. Review your local rules and consult a lawyer for commercial use.

How fresh is the data?

Every run pulls live data at execution time. Set up a schedule (daily/hourly/weekly cron) in the Apify console to keep your dataset current automatically.

Related actors in this category

Related tutorials

Open on Apify →