Structured web data, built for bulk

Datasets that run on the Apify platform and are made to be called again and again: a strict output schema, explicit error records, and a price per delivered row. Failed rows are never charged.

YouTube data

Transcripts, comments, channels and charts in bulk, for AI pipelines and content research. 10 datasets.

Google search and market data

Trends, autocomplete, news, jobs, shopping and local results, structured and repeatable. 15 datasets.

App, game and podcast data

App Store, Google Play, Steam and Apple Podcasts metadata, rankings and charts. 3 datasets.

Web and domain tooling

DNS, SSL, deliverability, tech stack, sitemaps, redirects and page extraction. 5 datasets.

E-commerce data

Shopify and WooCommerce catalogues and store reports. 2 datasets.

Company and financial data

SEC filings, finance quotes, careers pages, GitHub and Hacker News. 1 datasets.

Guides

How to pull this data, and where the free routes stop.

How to get YouTube transcripts in bulk

The free routes, the point where each one stops, and what it costs to pull transcripts for hundreds or thousands of videos.

Is there a Google Trends API?

What Google does and does not offer, why Trends numbers break when you compare them wrong, and the working routes.

How to export YouTube comments to CSV

Extensions, the official API and its quota, and how to pull comments for a whole channel without babysitting it.

How to scrape Google News results

The free RSS route, the Python route, and what changes once you need this every day without failures.

Is there a Google Jobs API?

Why the indexing API is not what most people are looking for, and how to read Google for Jobs results instead.