local-llm
SolidЛокальный LLM-стек на Mac (Apple Silicon, MLX) под приватность и запасной режим. NL-вход к установке/запуску/переключению моделей + слой суждения для мониторинга новых моделей. Тонкая обёртка над скриптами РП404, не замена.
Install
Quality Score: 82/100
Skill Content
Details
- Author
- TserenTserenov
- Repository
- TserenTserenov/FMT-exocortex-template
- Created
- 7 months ago
- Last Updated
- today
- Language
- Shell
- License
- MIT
Integrates with
Similar Skills
Semantically similar based on skill content — not just same category
llm_backends
Discover, audit and use local LLM servers (LM Studio, Ollama, LocalAI, vLLM, llama.cpp) on this machine or the network to offload work from the cloud and save tokens. Use when the user mentions local models, LM Studio, Ollama, "run it locally", token savings via local hardware, or wants to know what models their machine can run.
ai-local-model-ops
Runs local and self-hosted LLM workflows with Ollama, LM Studio, MLX, Open WebUI, llamafile, and adapters. Use when operating private model stacks.
local-llm-agent
Use the high-end LOCAL LLMs on this machine as a real agent/coding substrate (not just Claude). Dual RTX 5090 (64GB VRAM) + ollama already installed and serving. Best agentic-coding local model: qwen3-coder:30b. Covers how to pull, run, call (CLI / HTTP / OpenAI-compatible / function-calling), route through the AIOS provider harness, and use as a heterogeneous arm in absorption-probe. Per feedback_use_all_substrates_not_own_head — don't solve from one model.