# Managed web scraping: you define the data, we extract it.

Canonical URL: https://www.rastriq.es/en/managed-web-scraping/ · Spanish version: https://www.rastriq.es/scraping-a-medida/

> We build and run custom web extractions: process, anti-bot, delivery via Apify, API, CSV or S3, and compliance. 109 public actors as reference.

We design, run and maintain web extractions so your team receives structured data, not scripts to fix. From a single new source to a recurring pipeline with a stable schema.

QUICK ANSWER

**Quick answer:** Rastriq designs, runs and maintains custom web extractions on top of 109 public actors on the Apify Store. Data is delivered through Apify, the REST API, CSV or JSON files and S3 buckets, with no personal data and under the Rastriq data licence.

109

Public actors on Apify

4

Delivery channels

0

Personal data

Request a proposal → See the process

REAL SAMPLE

## What we already deliver, in real rows

Rastriq maintains 109 public actors on the Apify Store. That codebase and experience is the starting point for every custom project.

| Source | Public actor | Sample rows | Fields | Download |
|---|---|---|---|---|
| [Coches.net · car listings in Spain](https://apify.com/rastriq/cochesnet-spain) | cochesnet-spain | 58 | 55 | [CSV](https://www.rastriq.es/muestras/coches-net.csv) · [JSON](https://www.rastriq.es/muestras/coches-net.json) |
| [Wallapop · used cars](https://apify.com/rastriq/wallapop-cars-scraper) | wallapop-cars-scraper | 100 | 23 | [CSV](https://www.rastriq.es/muestras/wallapop-coches.csv) · [JSON](https://www.rastriq.es/muestras/wallapop-coches.json) |
| [Milanuncios · cars](https://apify.com/rastriq/milanuncios-scraper) | milanuncios-scraper | 100 | 30 | [CSV](https://www.rastriq.es/muestras/milanuncios.csv) · [JSON](https://www.rastriq.es/muestras/milanuncios.json) |
| [MachineryZone · used equipment in Europe](https://apify.com/rastriq/machineryzone-scraper) | machineryzone-scraper | 100 | 20 | [CSV](https://www.rastriq.es/muestras/machineryzone.csv) · [JSON](https://www.rastriq.es/muestras/machineryzone.json) |
| [Surplex · online industrial auctions](https://apify.com/rastriq/surplex-auction-scraper) | surplex-auction-scraper | 100 | 44 | [CSV](https://www.rastriq.es/muestras/surplex.csv) · [JSON](https://www.rastriq.es/muestras/surplex.json) |
| [BOE · official gazette entries](https://apify.com/rastriq/boe-scraper) | boe-scraper | 100 | 16 | [CSV](https://www.rastriq.es/muestras/boe.csv) · [JSON](https://www.rastriq.es/muestras/boe.json) |
| [Amazon · product reviews](https://apify.com/rastriq/amazon-reviews-scraper) | amazon-reviews-scraper | 24 | 23 | [CSV](https://www.rastriq.es/muestras/amazon-reviews.csv) · [JSON](https://www.rastriq.es/muestras/amazon-reviews.json) |
| [Google Ads Transparency Center · ads](https://apify.com/rastriq/google-ads-scraper) | google-ads-scraper | 100 | 15 | [CSV](https://www.rastriq.es/muestras/google-ads.csv) · [JSON](https://www.rastriq.es/muestras/google-ads.json) |
| [TikTok Creative Center · trending hashtags](https://apify.com/rastriq/tiktok-trends-scraper) | tiktok-trends-scraper | 50 | 12 | [CSV](https://www.rastriq.es/muestras/tiktok-trends.csv) · [JSON](https://www.rastriq.es/muestras/tiktok-trends.json) |
| [Europages · B2B directory](https://apify.com/rastriq/europages-scraper) | europages-scraper | 15 | 10 | [CSV](https://www.rastriq.es/muestras/europages.csv) · [JSON](https://www.rastriq.es/muestras/europages.json) |

Data from the current public samples. Each actor is published on the Apify Store and each sample can be downloaded without signing up.

### Example of structured output

| Country | Rank | Hashtag | Industry | Posts |
|---|---|---|---|---|
| ES | 1 | liveiseasy | News & Entertainment | 10500 |
| ES | 2 | octubre | News & Entertainment | 5600 |
| ES | 3 | labolanegra | News & Entertainment | 7500 |
| ES | 4 | lluvia | News & Entertainment | 3800 |
| ES | 5 | october | News & Entertainment | 2000 |

Five real rows from the TikTok Creative Center (hashtags) sample. It is an example of the structured data an extraction delivers.

PROCESS

## Five steps, from idea to pipeline

Every project follows the same path. The first sample arrives before you commit to the full build.

1 · Briefing

Source, fields, markets, volume and frequency. We clarify what you will use the data for and the format you need.

2 · Feasibility and compliance

We review how the information is published, which terms apply and whether the data is public. We rule out anything that requires third-party credentials or involves personal data.

3 · Prototype and sample

We deliver a real sample with the proposed schema so you can validate fields and quality before going to production.

4 · Production

We schedule the extraction, watch for errors and source changes, and apply normalization and QA on every run.

5 · Delivery and maintenance

The data arrives through the agreed channel and we maintain the actor when the source changes its structure.

ANTI-BOT

## How we handle anti-bot protections

We do not promise access to every website: we assess each source and choose the lightest technique that works reliably.

Real browsers and consistent sessions

When a source needs it, we use real browsers with consistent fingerprints and stable sessions instead of isolated requests.

Proxies and a conservative pace

We choose the proxy type by country and source, and cap concurrency so we never overload the origin site.

Internal API when available

If the site consumes a public API from its own frontend, we extract from there: it is faster, cheaper and more stable than rendering pages.

Change monitoring

Every run validates record counts and field shapes. If something changes, we catch it before it reaches your table.

DELIVERY

## Four delivery channels

We deliver where your team already works. If several sources must be unified, we add [data normalization](https://www.rastriq.es/en/data-normalization/) to the same pipeline.

Apify

A dataset in your Apify account, with scheduled runs and webhooks. You can rerun or extend the actor whenever you want.

Access: console and API

API

Query results through the Apify REST API with pagination and filters, and plug it into your ETL or your agents.

Format: JSON

CSV / JSON

Files ready to load into your data warehouse, with a field dictionary and the capture date of every row.

Formats: CSV, JSON

S3

Recurring delivery into an S3 bucket you own, with a folder structure by date and source.

Layout: source/date

COMPLIANCE

## Compliance: public data, licence and clear limits

**Public data only.** We extract information any visitor can see without logging in: prices, listings, product sheets, official publications. We do not collect personal data and we exclude phone numbers, emails and private individuals' names.

**Clear licence.** Data is delivered under Rastriq's [data licence](https://www.rastriq.es/en/data-license/), which describes the permitted use. Before we start we review the source terms and explain any limitation.

**No touching third-party systems.** We do not access private areas or bypass authentication. If a source requires credentials that are not yours, we do not extract it.

FAQ

## Frequently asked questions about managed scraping

### How long does a custom extraction take?

It depends on the source. For sites with an internal API or stable HTML, the first sample can be ready within a few days. Sources with anti-bot protection or many structural variants need more prototyping time.

### Who maintains the extraction if the website changes?

We do. Every run validates records and fields, and when the source changes its structure we update the actor. Maintenance is part of the recurring service.

### Can I keep the actor?

Yes, we can publish it in your Apify account or keep it in ours and deliver only the data. We agree this in the proposal, depending on whether you prefer code ownership or a managed service.

### How do I receive the data?

Through the Apify dataset with API access, as CSV or JSON files, or in an S3 bucket you own. We can also integrate webhooks to notify your system when each run finishes.

### Do you extract personal data?

No. Rastriq works with public product data and excludes phone numbers, emails, private individuals' names and other personal data. If a project requires them, we do not take it on.

FREE SAMPLE

## Tell us which source you need

Describe the site and the fields you are after. We reply with a feasibility assessment and a sample of the proposed schema.

- Feasibility assessment of the source
- Proposed schema and delivery channel
- No commitment and no personal data

Received. Meanwhile, download the real samples at [/en/samples/](https://www.rastriq.es/en/samples/).

CONTACT

## Is your source not on the Store?

Tell us what you need and we will assess it.

[Get in touch →](https://www.rastriq.es/en/#contacto) [See actors on Apify](https://apify.com/rastriq)

RELATED PAGES

## Keep exploring

[Buy datasets Catalogue with downloadable samples.](https://www.rastriq.es/en/buy-datasets/) [Data normalization A common schema across sources.](https://www.rastriq.es/en/data-normalization/) [Guide: build or buy web scraping How to decide between building and buying.](https://www.rastriq.es/en/guides/build-vs-buy-web-scraping/)
