MTCode Server
The desktop controller stores credentials, downloads selected models, configures the local server, starts and stops it, and publishes its TCP endpoint to authorized DirectLink users.
MTCode Server brings model selection, download, local inference, and secure publishing into one desktop workflow. Its built-in MTCode-LLM and MTCode-Diffusion servers are designed for people who want to get a useful local model running before learning a larger AI tool ecosystem.
A guided local AI setup
Instead of beginning with model formats, inference flags, and a long catalog, start with an application category. MTCode Server presents a preselected set of open-weight models hosted on Hugging Face, along with quantization and runtime choices intended to make the tradeoff between capability, speed, and GPU memory easier to understand.
Three pieces, one workflow
The desktop controller stores credentials, downloads selected models, configures the local server, starts and stops it, and publishes its TCP endpoint to authorized DirectLink users.
A local language and multimodal model server using the llama.cpp inference engine. It supports the text, coding, vision, and audio-oriented model choices exposed by MTCode Server.
A local image generation and editing server using the stable-diffusion.cpp engine. MTCode Server organizes supported workflows into generation, generation-and-editing, and inpainting categories.
Quick start for new users
The built-in path keeps the first setup inside one application. Advanced controls remain available, but they do not have to be the first thing a new user learns.
MTCode-LLM: select an application category, model, quantization, and context before starting the server.Click the image to enlarge.
MTCode-Diffusion: select a generation or editing workflow and a compatible image model.Click the image to enlarge.
Where it fits
Ollama, LM Studio, and established image-generation tools have broader ecosystems, integrations, catalogs, and community workflows. MTCode-LLM and MTCode-Diffusion are not presented as replacements for those tools. Their advantage is a smaller, guided path from selecting a useful open-weight model to running and sharing it from the MTCode Server GUI.
Use the built-in servers when application categories and a curated shortlist are more helpful than evaluating every runtime and model independently. They provide a practical starting configuration that can be refined later.
Keep it. MTCode Server can publish any TCP service, including an existing Ollama, LM Studio, or other compatible endpoint. DirectLink does not require the service to use an MTCode inference engine.
Models remain third-party software and content. Availability, hardware requirements, capabilities, and license terms vary by model. Review the Hugging Face model card and license before downloading or sharing access.
Local by design
Inference runs on the selected computer. DirectLink supplies authorized remote access to the server endpoint, while MTCode Portal gives users a stable local address and a familiar client workflow.
MTCode Server provides a direct route from local GPU hardware to a running language or image model—and an equally direct route for authorized users to reach it.