molla
Local inference server in pure Mojo. Speaks OpenAI, Anthropic, and MCP. Models are OCI artifacts. Same source runs on CPU, NVIDIA, AMD, and Apple GPUs. Apache-2.0 all the way down.
Local inference server in pure Mojo. Speaks OpenAI, Anthropic, and MCP. Models are OCI artifacts. Same source runs on CPU, NVIDIA, AMD, and Apple GPUs. Apache-2.0 all the way down.