halleriteKonstantin Dunas
On the Nature of the Swarm
or, why scaling inference ends in many agents
I’m a researcher at Prime Intellect. These days I spend a lot of time thinking about scaling multi-agent systems.
Much of my work is open source, including:
- the multi-agent abstraction in verifiers and prime-rl, for defining how agents interact and learn, from self-play to agentic judging.
- the algorithms layer in prime-rl, for defining training algorithms and mixing them across environments.
- renderers, a Python library for message–token conversion that preserves sampled tokens across turns.
Previously, I built Ludic, an exploration of the RL abstractions that have informed my subsequent work.
Κενὸς ὁ τάφος· τί οὖν ἡ καρδία σου ἔτι μεστὴ φόβου;✝
"The tomb is empty; why then is your heart still full of fear?"