맥에서 TurboFieldfare
Gemma 4 26B-A4B on Apple Silicon in ~2 GB of RAM · AI / LLMs
TurboFieldfare is a custom Swift 6.2 + Metal 4 runtime that runs Google's Gemma 4 26B-A4B (instruction-tuned) on any Apple Silicon Mac — including 8 GB machines — in roughly 2 GB of RAM, by streaming the mixture-of-experts weights from disk instead of holding them resident. Useful when you want a 26B-class local model on a Mac that can't hold one in memory; for other Gemma sizes use the Ollama gemma3:* tags in Model Manager instead, since this runtime serves only the one model. Built from source: git clone, swift build -c release, then run .build/release/TurboFieldfareMac (the ~14.3 GB model streams down on first launch). TurboFieldfareCLI gives instruction chat and raw completions with --temperature/--top-k/--max-new. An experimental loopback OpenAI-compatible Chat Completions server on http://127.0.0.1:8080/v1 adds streaming and function-tool support — add that in AI Models as a network endpoint to use it in the assistant. Apache-2.0. Note it builds unsigned third-party source rather than installing a notarized binary, and :8080 is also llama.cpp's default — change one if you run both.
FrontierStack으로 TurboFieldfare 실행하기
TurboFieldfare을(를) 설치하거나 연결한 뒤 FrontierStack에서 상태, 포트, 인증서를 확인하십시오. 해당되는 경우 같은 화면에서 방화벽 점검, 악성코드 감사, 백업으로 이동할 수 있습니다.
Mac에서 모두 운영하십시오.
FrontierStack은 이 Mac과 연결된 서버에서 서비스를 설치, 모니터링, 보호하며 — 암호와 키는 어떤 AI에게도 보이지 않게 지킵니다.
FrontierStack 다운로드Apple 공증 완료 · 안전 & 보안 · macOS 13 Ventura 이상
