The Wullup Crawler
Last Updated: 23.09.2026
This page is for website operators who have come across our crawler in their server logs.
This notice is available in English and German. In case of discrepancies, the German version prevails.
1. What the Crawler Does
The Wullup crawler is the event crawler of Wullup GmbH. It reads only the public event listings on websites of venues and organizers and extracts factual event details from them — title, date, time, venue, and price — which we show in the Wullup app and on wullup.com.
Which details we take, what we do not take, how we handle personal data, and how long we keep what are explained in our Event Sources notice.
2. How to Identify It
- User-Agent:
Wullup/1.0 (+https://wullup.com/bot) - Origin: Requests come from Wullup's own servers — no rotating IP addresses and no proxies
- Pages that show content only with JavaScript: We may have such a page rendered once by a hosted headless browser (Apify). That request sends the same User-Agent
3. What the Crawler Respects
3.1 robots.txt
The crawler respects robots.txt under the Robots Exclusion Standard (RFC 9309), including a waiting time stated there (Crawl-delay). It follows the group User-agent: Wullup, otherwise the group User-agent: *.
To block the crawler completely:
User-agent: Wullup
Disallow: /
3.2 Text and Data Mining Reservations (§ 44b(3) UrhG)
Any one of the following reservations is enough for the crawler to stop reading the website and for the events already taken from that website to be removed:
- TDMRep:
/.well-known/tdmrep.json, thetdm-reservationHTTP header or meta tag noai/noimageaiinX-Robots-Tagor in the robots meta tag- IETF AIPREF: a
Content-Usagepreference in the HTTP header or as a line in robots.txt - A written reservation of use or a ban on automated analysis in the website's terms of use, terms and conditions, or imprint
Images are checked the same way.
3.3 No Circumvention
The crawler never logs in and does not solve captchas. If a website answers with a bot or captcha challenge, it stops the whole run for that website.
4. Pace
Each website is retrieved roughly every three and a half days, with at least 0.5 seconds between two requests — or more if your Crawl-delay asks for it.
5. Exception: The Rights Holder's Permission
Where a venue has registered its own website in Wullup Backstage, consented to the taking, and verified that the website is theirs, the rights holder's permission is in place. For that website, its text and data mining reservations and bans in its terms of use no longer stop the crawler. A Disallow in robots.txt is still respected.
6. Contact and Opt-Out
You do not want us to take your events? Use our opt-out form or write to [email protected]. For questions about the crawler, you can also reach us at [email protected].