홈 › 서비스 › MLX

맥에서 MLX

Apple-silicon model serving — the MLX-LM CLI server, the oMLX menu-bar app, or Inco AI's Splash engine · AI / LLMs

MLX runs large language models natively on Apple Silicon — unified M-series memory, often faster and lighter than llama.cpp on a Mac, with models from the mlx-community org on Hugging Face. FrontierStack manages both ways to serve them, in this one pane: MLX-LM (pip install mlx-lm, then mlx_lm.server --model … --port 8080 — a plain OpenAI-compatible API) and oMLX (a native menu-bar app on :8000 that is a drop-in replacement for the OpenAI *and* Anthropic APIs, serving LLM/VLM/OCR/embedding/reranker models with continuous batching, a two-tier KV cache, and a web model-manager at /admin; install the .dmg from omlx.ai or brew install omlx). Point LibreChat, OpenCode, Claude Code or the Model Costs pane at either; all three appear in the AI-harness model picker when running. The pane also runs Splash (Inco AI's open-source Apple-silicon engine, brew install incoai/tap/splash; OpenAI *and* Anthropic APIs; M3+ Mac, macOS 26.4+, 36 GB+ memory) — not MLX-format models, but the same kind of local server; FrontierStack starts it on :8100 so it does not collide with oMLX on :8000.

FrontierStack으로 MLX 실행하기

MLX을(를) 설치하거나 연결한 뒤 FrontierStack에서 상태, 포트, 인증서를 확인하십시오. 해당되는 경우 같은 화면에서 방화벽 점검, 악성코드 감사, 백업으로 이동할 수 있습니다.

Mac에서 모두 운영하십시오.

FrontierStack은 이 Mac과 연결된 서버에서 서비스를 설치, 모니터링, 보호하며 — 암호와 키는 어떤 AI에게도 보이지 않게 지킵니다.

FrontierStack 다운로드

Apple 공증 완료 · 안전 & 보안 · macOS 13 Ventura 이상

관련 AI / LLMs