About Kabuuga
A UK data collection service built around one idea: research deserves web data that is open, documented and collected with care.
Why we exist
The web holds a huge record of regional life: public notices, local news, organisations, events, services and more. Turning that into a dependable research dataset takes engineering, patience and a clear ethical stance.
Kabuuga does that groundwork so researchers can focus on questions, not scrapers. We collect only what is publicly available, keep records of how we did it, and hand over data your peers can scrutinise.
Our mission
To make publicly available regional web data easy to use for academic and research work, without cutting corners on legality, ethics or quality.
What we believe
Reproducibility
If a result cannot be traced to its data, it cannot be trusted. Datasheets and logs are part of every delivery.
Respect for sources
Website owners run real services. We stay light on their servers and follow their published rules.
Privacy by default
We design collections to avoid personal data unless there is a clear, lawful and approved need.
Honest limits
Web data is incomplete and changes. We state coverage gaps plainly instead of hiding them.
What we do not do
No private or gated content
We do not log in, share credentials, or get around paywalls, CAPTCHAs or technical blocks.
No surveillance or profiling
We do not track individuals, build profiles of people or support targeting.
No resale of raw scrapes
Datasets are prepared for the research purpose agreed with you, with licence notes included.