Models
Open weights without lock-in
- Ollama models available in chat
- Hugging Face repositories running through MLX or llama.cpp
- Qwen and Gemma parameters adjustable during chat
- Context and tokens measured against the real limit
AI infrastructure
Choose local or online models and see the real state of the Mac. T-Chat shows what is ready, what uses resources and which limits matter.
A local control layer
Ollama, Hugging Face models, ComfyUI and specialised runtimes appear in the interface when they are genuinely ready.
Models
Companion
Expansion
Open optimisation
The experience is optimised for the Qwen family and Google Gemma without locking the platform to one provider.
Temperature, top-p, top-k, min-p, seed and context can be adjusted for models that expose them.
Charts and status come from observed processes, memory and requests rather than decorative animation.
Ready local models are registered through Companion. TimesFM 3 is activated separately as a forecasting tool for personalities, chats, meetings, missions and Automata.
Supported MoE models keep experts on fast storage and load the required ones into memory. It is slower than RAM and is not universal swap.
Build your AI perimeter
Let's assess the machine, runtimes and requirements for a sustainable local configuration.