Popular repositories Loading
-
llm-inference-optimization-lab
llm-inference-optimization-lab PublicReproducible llama.cpp CPU inference profiling and a deterministic LLM serving simulator with continuous batching, KV cache, prefix caching, and workload-driven latency analysis.
-
llm-engineering-platform
llm-engineering-platform PublicA production-oriented LLM engineering platform with OpenAI-compatible serving, streaming, observability, evaluation, reproducible experiments, and deterministic RAG.
Python 17
-
omiv-kimi-k3-validation
omiv-kimi-k3-validation PublicEvidence-backed structural validation of Kimi K3 UD-IQ1_M and UD-Q4_K_XL split GGUF releases using OMIV.
Shell 14
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
