About Kabuuga

A UK data collection service built around one idea: research deserves web data that is open, documented and collected with care.

Why we exist

The web holds a huge record of regional life: public notices, local news, organisations, events, services and more. Turning that into a dependable research dataset takes engineering, patience and a clear ethical stance.


Kabuuga does that groundwork so researchers can focus on questions, not scrapers. We collect only what is publicly available, keep records of how we did it, and hand over data your peers can scrutinise.

Our mission

To make publicly available regional web data easy to use for academic and research work, without cutting corners on legality, ethics or quality.

What we believe

🔍

Reproducibility

If a result cannot be traced to its data, it cannot be trusted. Datasheets and logs are part of every delivery.

🤝

Respect for sources

Website owners run real services. We stay light on their servers and follow their published rules.

🔒

Privacy by default

We design collections to avoid personal data unless there is a clear, lawful and approved need.

🧭

Honest limits

Web data is incomplete and changes. We state coverage gaps plainly instead of hiding them.

What we do not do

No private or gated content

We do not log in, share credentials, or get around paywalls, CAPTCHAs or technical blocks.

No surveillance or profiling

We do not track individuals, build profiles of people or support targeting.

No resale of raw scrapes

Datasets are prepared for the research purpose agreed with you, with licence notes included.

Not legal advice. We aim to follow good practice, but each institution remains responsible for confirming that a dataset suits its own ethics approvals and legal obligations.

Let us talk about your project

A short message is enough to get started.

Get in touch