Runs LLM inference on CPU, Apple Silicon, and consumer GPUs without NVIDIA hardware. Use for edge deployment, M1/M2/M3 Macs, AMD/Intel GPUs, or when CUDA is unavailable. Supports GGUF quantization (1.
12-inference-serving/llama-cpp/SKILL.md
Use this skill when the user asks to save, remember, recall, or organize memories. Triggers on: 'remember this', 'save t...
CLI tool for configuring and monitoring Claude Code
指导Claude按照二哥的风格撰写求职类文章,包括公司薪资爆料、年终奖盘点、求职攻略、offer选择建议等内容。