Popular repositories Loading
-
pxq_llama.cpp
pxq_llama.cpp PublicPXQ: PXA-native low-bit MoE quants (2/3/4-bit, E16-row scales) + fused CUDA kernels for Pascal/Volta — run a real 35B on a salvaged 12-16GB card. Fork of ik_llama.cpp.
-
awesome-local-ai
awesome-local-ai PublicForked from janhq/awesome-local-ai
An awesome repository of local AI tools
-
PXA_llama
PXA_llama PublicRun modern hybrid/MoE LLMs correctly and fast on cheap old Tesla P100 / GTX 1080 Ti cards. Fork of ik_llama.cpp: clean concurrent (np>1) Gated-DeltaNet hybrid decoding + Pascal sm_60 FP16 build tun…
Shell
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.

