The NanoScrape Compliance Scanner Scraper is an Apify actor that extracts structured web data from Compliance Scanner. It returns 17 fields per result including url, status_code, site_language, has_impressum, has_privacy_policy, priced at $3/1k url scanned, with no API key or monthly subscription required.
Scan any website for legal compliance markers: Impressum / imprint, privacy policy, cookie-consent banner. Returns a HIGH / MEDIUM / LOW risk flag tuned for DACH GDPR / Abmahnung exposure. Pay-per-result.
| Field | Type | Example | Group |
|---|---|---|---|
company_id | string | null | (populated when data available) | Core |
url | string | (populated when data available) | Core |
final_url | string | null | (populated when data available) | Core |
impressum_url | string | null | (populated when data available) | Core |
privacy_policy_url | string | null | (populated when data available) | Core |
status_code | integer | null | (populated when data available) | Other |
site_language | string | null | (populated when data available) | Other |
has_impressum | boolean | null | (populated when data available) | Other |
has_privacy_policy | boolean | null | (populated when data available) | Other |
cmp_detected | boolean | null | (populated when data available) | Other |
cookie_consent_status | string | (populated when data available) | Other |
consent_tool | string | null | (populated when data available) | Other |
sets_tracking_cookies | boolean | null | (populated when data available) | Other |
risk_flag | string | null | (populated when data available) | Other |
risk_reasons | array | null | (populated when data available) | Other |
probed_at | string | null | (populated when data available) | Other |
error | string | null | (populated when data available) | Other |
{
"urls": [
"https://www.deutschebank.de/",
"https://www.microsoft.com/",
"https://www.example.com/"
],
"probeTimeoutMs": 10000,
"respectRobotsTxt": true,
"concurrency": 3,
"proxyConfiguration": {
"useApifyProxy": false
}
}Priced at $0.0030 per url scanned. 10 url scanneds ≈ $0.03, 100 ≈ $0.30, 1,000 ≈ $3.00. The first runs land inside Apify's $5/mo free-tier credit — pay only for what you extract, no monthly subscription.
Click Open on Apify above to run compliance-scanner in your browser - no code, no install. In the Apify console you get:
To run this actor from your own code you need an Apify API token. Get one in about a minute:
Free tier: Apify credits your account with $5 of platform usage every month, no credit card required. Enough to test any actor meaningfully - at $0.001 per result on typical scrapers, that is roughly 5,000 results for free every month.
Run this actor from any language via the Apify REST API. Replace YOUR_TOKEN with your API token and adapt the input JSON to your needs.
curl -X POST "https://api.apify.com/v2/acts/santamaria-automations~compliance-scanner/run-sync-get-dataset-items?token=YOUR_TOKEN" \
-H "Content-Type: application/json" \
-d '{}'// npm install apify-client
import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: 'YOUR_TOKEN' });
const run = await client.actor('santamaria-automations/compliance-scanner').call({});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);# pip install apify-client
from apify_client import ApifyClient
client = ApifyClient('YOUR_TOKEN')
run = client.actor('santamaria-automations/compliance-scanner').call(run_input={})
items = list(client.dataset(run['defaultDatasetId']).iterate_items())
print(items)// dotnet add package Apify.Client
using Apify.Client;
var client = new ApifyClient("YOUR_TOKEN");
var run = await client.Actor("santamaria-automations/compliance-scanner").CallAsync(new { });
var items = await client.Dataset(run.DefaultDatasetId).ListItemsAsync();// Maven: com.apify:apify-client
import com.apify.client.ApifyClient;
ApifyClient client = new ApifyClient("YOUR_TOKEN");
ActorRun run = client.actor("santamaria-automations/compliance-scanner").call(Map.of());
List<Map<String,Object>> items = client.dataset(run.getDefaultDatasetId()).listItems();This actor is available on the Apify MCP server, so you can drive it from any MCP-compatible AI client - Claude Desktop, Claude.ai, Cursor, VS Code, LangChain, LlamaIndex, or a custom agent - without writing any code.
https://mcp.apify.com?tools=santamaria-automations/compliance-scannerExample prompt once connected:
"Use <code>compliance-scanner</code> to run a scrape on my target list and give me the results as a table."
Clients that support dynamic tool discovery (Claude.ai, VS Code) receive the full input schema automatically via add-actor.
Trigger this actor from your favorite automation tool. Every platform below can call the Apify API in a few clicks - no code required.
@apify/n8n-nodes-apify).santamaria-automations/compliance-scanner.https://api.apify.com/v2/acts/santamaria-automations~compliance-scanner/run-sync-get-dataset-items?token=YOUR_TOKEN.application/json.https://api.apify.com/v2/acts/santamaria-automations~compliance-scanner/run-sync-get-dataset-items?token=YOUR_TOKEN, method POST, body raw JSON.https://api.apify.com/v2/acts/santamaria-automations~compliance-scanner/run-sync-get-dataset-items?token=YOUR_TOKEN.steps.http.$return_value in subsequent steps.https://api.apify.com/v2/acts/santamaria-automations~compliance-scanner/run-sync-get-dataset-items?token=YOUR_TOKEN, body type JSON.$3.00 / 1,000 url scanned, billed per result on Apify. Apify's $5/month free tier covers the first runs. No monthly subscription — you only pay for what you extract.
17 fields per result. Full field catalog is listed in the 'What you can scrape' table above — includes core identifiers, descriptive text, dates, and any category-specific data.
Scraping publicly available data is generally permitted in most jurisdictions, but GDPR applies to any personal data you collect. Review your local rules and consult a lawyer for commercial use.
Every run pulls live data at execution time. Set up a schedule (daily/hourly/weekly cron) in the Apify console to keep your dataset current automatically.