Built for Tribunus Desktop — the AI coding agent. Powered by Tribunus Compute — the portable inference engine with compile-time architecture.
Your AI coding agent. Runs locally. Plugin system. Multi-agent. Any model, any backend. Electron desktop app for macOS, Linux, Windows.
The inference engine inside Tribunus Desktop. Compile-time architecture — every decision frozen before runtime. Numerical oracle validates every kernel. Multi-backend from day one.
Architecture decision records that define the Tribunus Compute engine — each one generated, verified, and source-backed.
Defines the six runtime pipelines — token intake, prefill, decode, KV cache management, speculative decoding, and output streaming — their scheduling guarantees, memory budgets, and the handoff contracts between adjacent stages. Each pipeline is a separate phase in the compile-time planner.