中文
Agent infrastructure

ik_llama.cpp

ikawrakow/ik_llama.cpp

Fork of llama.cpp focused on local LLM inference, adding state-of-the-art quantization formats and CPU performance optimizations for faster execution on consumer hardware.