Redditor Cherlokoms reports that macOS 27 exposes Apple's Foundation Models locally through the native `fm chat` command, potentially making on-device inference accessible directly from a Mac terminal.
state: resolvedheat: lowuncertainty: mediumconvergesscott: lowlocal-inference apple-foundation-models on-device-aiAppleCherlokoms
Surfaced 2026-09-21T15:41:40Z — Apple Foundation Models: local AI natively on MacOS 27 — Cross-platform attention and the macOS model-download workaround strengthen the broader evidence that Apple is distributing on-device models, but they still do not independently verify the native `fm chat` interface. The loud spread reading warrants attention without promoting the specific claim to corroborated.
What is this?
The supplied secondary reports describe Apple introducing a native `fm` command-line tool with macOS 27 to prompt its on-device Foundation Models from the terminal, alongside a Python SDK for scripting. The third-party `fmx` repository describes providing a similar chat/respond/schema workflow on macOS 26 while anticipating the native command. No direct Apple documentation or Cherlokoms post is supplied, and the snippets differ on shipping status and installation requirements, so the precise availability and native `fm chat` command remain incompletely verified. Claims of a built-in `fm serve` OpenAI-compatible server are explicitly disputed by ChatForest’s correction and should not be treated as established.
Why it matters to Scott
Apple’s reported native terminal inference interface converges with Scott’s existing `ask` local-model path and makes a potential additional backend worth evaluating for his LiteLLM-routed tooling; it does not establish an equivalent shell-executing agent or a drop-in OpenAI-compatible endpoint. Availability remains incompletely verified, and the radar’s related `apple-silent-on-device-model-changes` and `deforget-apple-model-drift` pages raise reproducibility questions for that evaluation rather than already covering this CLI development.
dev:project.askdev:technology.litellmdev:technology.ollamaradar:concept.local-inferenceradar:concept.on-device-airadar:apple-silent-on-device-model-changesradar:deforget-apple-model-drift
queries asked of Scott's wikis
- local inference economics and cloud dependency
- OS-bundled AI versus model choice and sovereignty
- terminal AI scripting and coding-agent harness integration
- Apple Silicon local model deployment projects
- private document processing and on-device knowledge systems
- OpenAI-compatible adapters and inference provider portability
Measured heat
no measured readings yet — the hourly heat pass fills this in
How the heat travelled
Evidence (3) — ⭐ canonical anchor
Interpretation history
2026-09-28T05:27:03Z
The new apple-llm wrapper targets macOS 26+, showing the built-in AFM was already scriptable before 27 — so the only distinctive part of Cherlokoms's claim, a native `fm chat` CLI, is now a moot technicality: access was never the bottleneck for Scott's tooling, model quality is, and that verdict is settled negative (worse than Gemma E4B, ~8k ctx, refusals, temp-0 repetition, cold starts). With the thread flatlined for a week (0.17 pts/h vs 81.7 peak), no periphery expansion, and the live reproducibility themes carried by sibling cases, the episode ends window-closed rather than lingering as a dormant watcher.
2026-09-28T05:23:06Z
evidence attached: reddit.post.1ws5l5p — Third-party apple-llm library plus reported AFM limitations (temp-0 repetition, cold-start latency) directly bear on whether macOS's built-in local LLM becomes practically usable.
2026-09-24T12:51:09Z
The first first-hand test report lands negative — the on-device model is worse than Gemma E4B with ~8k context and heavy refusals — which removes the main practical reason to chase `fm chat` as an `ask`/LiteLLM backend, so Scott-relevance drops to low. The thread has flatlined (0 pts/h vs a 78/h peak, 11th percentile) with no new platforms or implementations, so heat cools to low; the loud-spread alert already fired on Sep 21 and the access claim itself remains unverified at watching.
2026-09-21T15:41:40Z
magnitude valve eligible (multi-platform, top-decile engagement) and never alerted; deterministic escalation to deliver
2026-09-21T15:25:00Z
evidence attached: hn.story.49787535 — The workaround materially contextualizes macOS Foundation Model downloads and local-storage requirements.
2026-09-15T17:33:17Z
grounded: converges/medium — Apple’s reported native terminal inference interface converges with Scott’s existing `ask` local-model path and makes a potential additional backend worth evalu
2026-09-15T17:26:57Z
case created — The specific access claim is checkable and distinct from the existing Apple model-drift and private Siri integration cases, but currently rests on one community report.
Decision trace
- 09-28 15:27resolveThe new apple-llm wrapper targets macOS 26+, showing the built-in AFM was already scriptable before 27 — so the only distinctive part of Cherlokoms's claim, a native `fm chat` CLI, is now a moot
- 09-28 15:23attachThird-party apple-llm library plus reported AFM limitations (temp-0 repetition, cold-start latency) directly bear on whether macOS's built-in local LLM becomes practically usable.
- 09-28 15:23propose_attachThird-party apple-llm library plus reported AFM limitations (temp-0 repetition, cold-start latency) directly bear on whether macOS's built-in local LLM becomes practically usable.
- 09-26 06:34review_screenjev screen: no material development (noul=0.10)
- 09-24 22:51repriceThe first first-hand test report lands negative — the on-device model is worse than Gemma E4B with ~8k context and heavy refusals — which removes the main practical reason to chase `fm chat` as an `as
- 09-24 22:49review_screenNew first-hand comment from CozyPinetree reports actually testing the on-device model: worse than Gemma E4B, only 8k context, and heavy refusal behavior. This is a new implementation/experience result
- 09-24 22:49review_screenjev screen borderline (noul=0.59) — luna review
- 09-22 01:41repriceCross-platform attention and the macOS model-download workaround strengthen the broader evidence that Apple is distributing on-device models, but they still do not independently verify the native `fm
- 09-22 01:41pushApple Foundation Models: local AI natively on MacOS 27 — Cross-platform attention and the macOS model-download workaround strengthen the broader evidence that Apple is distributing on-device models, b
- 09-22 01:41alert_routeApple Foundation Models: local AI natively on MacOS 27 — Cross-platform attention and the macOS model-download workaround strengthen the broader evidence that Apple is distributing on-device models, b
- 09-22 01:25attachThe workaround materially contextualizes macOS Foundation Model downloads and local-storage requirements.
- 09-22 01:23propose_attachThe workaround materially contextualizes macOS Foundation Model downloads and local-storage requirements.
- 09-19 21:27review_screenThe changes remove an opinion about model quality and add a speculative converter comment, without providing new evidence about terminal access or local Foundation Models availability.
- 09-18 18:23review_screenThe added comment is a subjective quality report about model performance and refusals; it does not confirm, refute, or materially change the claim about terminal access.
- 09-16 10:33review_screenThe change adds only an incomplete comment reiterating that the models target simple on-device use cases, with no new evidence about terminal access or implementation.
- 09-16 10:21sensor_dirtycomment_update
- 09-16 03:33groundApple’s reported native terminal inference interface converges with Scott’s existing `ask` local-model path and makes a potential additional backend worth evaluating for his LiteLLM-routed tooling; it
- 09-16 03:26createThe specific access claim is checkable and distinct from the existing Apple model-drift and private Siri integration cases, but currently rests on one community report.