Local Inference Optimizer
skill-forgetmeai-local-inference-optimizer-skill-local-inference-optimizer-skill · by ForgetMeAI
Choose and tune a local/production LLM inference engine for the user's hardware, model, and workload; set up uv+venv; select kernels/quantization; tune batching, KV cache, flags, and launch commands.