Discoverability
Store classification
Categories
Store labels
No public values available.
Catalog ID Emy31Y0dszCLXycar
Public Apify Actor
parseforge/poland-open-data-scraper
Scrape dane.gov.pl, Poland's national open data portal: 26,922 datasets, 1.6M resources, 7,471 institutions, showcases, change history and table rows over the public JSON:API.
Active users / 30d
1
Total users
2
Runs / 30d
18
Total runs
36
Rating
—
No reviews
Bookmarks
0
Usage history
Latest captured Active users / 30d
1
0 since 28 Aug
19 of 30 UTC days captured
Latest captured Runs / 30d
18
+17 since 28 Aug
19 of 30 UTC days captured
1
Users / 7d
Discoverability
Categories
Store labels
No public values available.
Monetization
Pay per event
19 charge events configured
Actor Start
apify-actor-start
Charged when the Actor starts running. Number of events charged depends on Actor memory (one event per GB, minimum one event).
$0.02
Catalogue page scanned
catalogue-page
One page of up to 50 records read from a dane.gov.pl index. This is the fixed cost of paging through the catalogue, spread across every row that page yields.
1
Users / 30d
1
Users / 90d
$0.004
Dataset
dataset-row
One dane.gov.pl dataset with its full 49-field record: Polish title and description in text and HTML, portal and API URLs, publishing institution with its type, city and website, EU DCAT themes with their codes, the legacy portal category, keywords, geographic regions, every file format and delivery type it offers, resource and showcase counts, licence name, code, plain-English description and every reuse condition, update frequency, openness score, the high value, EU high value, dynamic, research and promoted flags, view and download counts, the external catalogue it was harvested from, supplementary documents, and the created, modified, verified and resource-modified timestamps.
$0.007
Resource
resource-row
One distribution: the downloadable file, API or service inside a dataset, with its format, media type, byte size, direct download URL, CSV and JSON-LD conversions, openness score, language, data date, the protected-data and high-value flags, the special-sign legend the publisher used for missing cells, and the dataset and institution it belongs to.
$0.003
Institution
institution-row
One of the 7,476 publishing bodies with its REGON statistical number, institution type, postal address, city, email, telephone, fax, ePUAP inbox, electronic delivery address, website, dataset and resource counts, and the external catalogues it is harvested from.
$0.008
Showcase
showcase-row
One of the 119 public services, websites and apps built on dane.gov.pl data, with its author, external URL, licence type, platform, mobile and desktop store links, keywords, screenshots and how many datasets it reuses.
$0.006
Change history entry
history-row
One entry from the portal audit log: the action, the timestamp, the editor account, the record it touched, which fields changed and the full before-and-after difference the portal recorded.
$0.003
Broken link
broken-link-row
One row from the portal's own dead-link report: the institution, the dataset, the resource and the URL that stopped resolving, with the date the report was last rebuilt.
$0.004
News item
news-row
One portal news article with its title, full text, keywords, author and publication date. News is only reachable through the cross-model search index, never through a listing endpoint.
$0.004
Knowledge base article
knowledge-base-row
One knowledge base article: the portal guidance, reports and training material for data publishers, with its category, keywords and full text.
$0.004
Search result
search-row
One hit from the cross-model search index, which is the only place the portal filters on licence, update frequency and date, and the only way to query datasets, resources, institutions, showcases, news and knowledge base in a single pass.
$0.003
Table row
tabular-row
One real row of data out of a resource's own table, with the publisher's column headers restored instead of the API's col1, col2, col3, plus the row number and the date that row was last refreshed.
$0.002
Catalogue breakdown
facet-row
One aggregation bucket: a format, EU theme, portal category, delivery type, openness score or publisher with its dataset count and its share of the filtered catalogue. This is the shape of the catalogue rather than its contents.
$0.002
Files of a dataset
dataset-resources
Opt in. Every file inside one dataset attached to its row, each with its format, byte size, openness score, download URL, data date and visualisation types. Billed only when the dataset actually returned files.
$0.004
Apps reusing a dataset
dataset-showcases
Opt in. The public services and apps that reuse one dataset, attached to its row with their authors, URLs and licence types. Billed only when the dataset actually has one.
$0.004
Change history of a dataset
dataset-history
Opt in. The portal audit trail for one dataset attached to its row: every update with its timestamp, the editor account and the fields that changed, plus the total change count. Billed only when the audit log had entries.
$0.004
Full publisher profile
institution-profile
Opt in. The publishing body's complete record attached to the row: REGON, postal address, email, telephone, fax, ePUAP inbox, electronic delivery address, website, dataset and resource counts and harvest sources. Billed only when the profile came back.
$0.005
Table preview
tabular-preview
Opt in. The column schema with the publisher's real header names and types, the total row count, and the first rows of a resource's table attached to its row. Only tabulated resources have any, and the rest are not billed.
$0.004
Live file check
file-probe
Opt in. One HEAD request against the actual download URL recording the HTTP status, content type and byte size, which is how you find files that stopped resolving before the portal notices. Billed only when the check completed.
$0.003
Loading public Actor documentation
Reading the current public Store definition and reviews.
More from this builder
Tennis Abstract Scraper - Match History & Stats
781 active users · 13.6K runs / 30d
Reddit Scraper - Posts, Subreddit & Search Data API
145 active users · 3K runs / 30d
TennisExplorer Match Results Scraper
145 active users · 500 runs / 30d
TikTok Hashtag Scraper - View Counts & Top Posts
78 active users · 946 runs / 30d