Models and datasets, prepared once into Semurg's on-disk form.
Big language models and large public datasets are the slow, heavy part. We do that work once: we turn them into Semurg's own fast on-disk form so your node reads them directly, with no convert step and no import on your side. Delivery is license-gated and inside your control, and each package verifies itself as it lands. Nothing about your own data ever leaves your node.
The intelligence surface is the canonical home for how models run on Semurg: intelligence.semurg.io. It states plainly what is serving today and what is still building. This page is the catalogue of what you can bring in, tagged honestly, part by part.
New to Semurg? → Start here, get a node running first. This page is the world you bring into it.
Transform once. Deliver to your node.
No jargon. Here is the whole idea in seven plain steps.
Models.
A model is something you run to generate: give it a prompt and it produces text and answers. These stream in and run on your own CPU, with no graphics card needed. Each card says plainly what it is, roughly how big it is, and what your machine needs.
Gemma 4 E4B
Qwen3
Kimi K3
Your models and extensions
Datasets.
A dataset is not a program, it is data you explore: real records you load, see as a live graph, and query. A model generates; a dataset is what you look at and traverse. Same streaming, same free check on arrival.
Wikidata
Other public datasets
What works today, and what is next.
We would rather undersell than overclaim. Here is exactly where each part stands.