Research
Research is the argument. Infrastructure is the proof.
Three programmes. Every position below either ships in a system or gets retired.

R.01
LUGHA
Sovereign language models
Models built for the languages institutions actually operate in. A decision-maker briefed by a system that cannot reliably work in their working language is constrained by the tool, so linguistic fluency is treated as a first-order engineering requirement rather than a translation layer added afterwards. Current artifact: KW5-Lite, a 109.5M parameter Kiswahili foundation model trained from scratch on 2.1B tokens, with instruction-tuned variant available.
- Trained from scratch on 2.1B Swahili tokens
- Holds a short exchange in standard-register Kiswahili on modest hardware
- Degrades after 7 to 10 turns, and is not for consequential work
- Open weights, Apache 2.0, so the claims can be checked
R.02
UAMUZI
Decision simulation
Deterministic systems that let institutions rehearse consequential decisions and delegate execution safely. Infrastructure, procurement and public-health decisions unfold over decades, so the systems that guide them have to produce the same outputs from the same inputs: auditable, reproducible, explainable. Rehearsal without determinism is theatre. Delegation without constraints is reckless.
Feeds Wallgarden
- Same inputs produce the same outputs: determinism as a product principle
- Decision simulation: what happens if, answered before money moves
- Organizational memory: outcomes compound instead of being relearned
R.03
MAANA
Semantic grounding
Meaning treated as infrastructure. One machine-verifiable, governed definition of every concept, so systems integrate cleanly and AI stays grounded in fact instead of hallucinating across mismatched definitions. Every consequential output must be explainable back to source facts and rules, with defined behavior under disagreement.
Feeds Matta · Cordon
- OWL 2 DL and SHACL with strict, non-overlapping mandates
- A deterministic truth hierarchy and a temporal graph
- AI components consume only authoritative facts; probabilistic ones are labeled
The register
What's on the bench.
Nothing here yet.
Published notes
Nothing published yet.
Standards
- PDPA 2022Data Protection, Tanzania
- e-GAe-Government Interoperability
- BoT Cyber RiskFinancial Data Isolation
- Ed25519 · MerkleTamper-Evident Logging
- OWL 2 DL · SHACLSemantic Governance
- Apache 2.0Open Weights: KW5-Lite
Open releases