# LAIONDownloader

> LAION's downloader, used to materialise the image and text datasets the non-profit publishes for machine-learning research. LAION's own FAQ is the source for its robots.txt position.

| field | value |
|---|---|
| operator | LAION / img2dataset |
| category | Corpus and dataset builders |
| robots.txt token | `LAIONDownloader` |
| user-agent contains | `LAIONDownloader` |
| robots.txt | not governed by robots.txt (user-initiated, by operator policy) |
| verify by | no published verification method |
| published IP ranges | none published |
| prefixes mirrored | 0 IPv4 / 0 IPv6 |
| operator docs | https://laion.ai/faq/ |
| last reviewed | 2026-09-01 |

## What blocking it costs you

Your media is skipped when an open research dataset is built from URL lists. Once a dataset is published, a later block does not remove you from it.

## Full user-agent string

```
LAIONDownloader
```

## Block it

```
User-agent: LAIONDownloader
Disallow: /
```

## Allow it

```
User-agent: LAIONDownloader
Allow: /
```

JSON: https://www.pathwren.workers.dev/crawler/laiondownloader.json · index: https://www.pathwren.workers.dev/llms.txt
