I build and improve custom toolsets for different open sourced and local LLM solutions that no one else has thought of yet!
Popular repositories Loading
-
GLM-5.3-Flash-EXL3-K2-DGX-Spark-recipe
GLM-5.3-Flash-EXL3-K2-DGX-Spark-recipe PublicvLLM recipe: GLM-5.3-Flash EXL3 K2 on one DGX Spark GB10. Native MTP k=2. Measured tok/s.
-
-
hermes-agentic-bench
hermes-agentic-bench PublicAgentic test batteries for local models via Hermes Agent - scripted tools + real Hermes CLI sessions
Python 14
-
DeepSeek-V4-Flash-Vision-EXL3-MixedK-DGX-Spark-recipe
DeepSeek-V4-Flash-Vision-EXL3-MixedK-DGX-Spark-recipe PublicServe DeepSeek-V4-Flash-Vision EXL3 MixedK (full 256 experts) on one NVIDIA DGX Spark with vLLM — recipe, loader patches, and receipts
-
Qwen3.8-Flash-Next-EXL3-DGX-Spark-recipe
Qwen3.8-Flash-Next-EXL3-DGX-Spark-recipe PublicServe Qwen3.8-Flash-Next (turboderp ExLlamaV3 pack) on one NVIDIA DGX Spark with vLLM + vllm-exl3
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.




