When setting up local LLM inference without cloud APIs. When running GGUF models locally. When needing OpenAI-compatible API from a local model. When building offline/air-gapped AI tools. When trouble
community/jamie-bitflight-claude-skills/llamafile/SKILL.md(main)