Skip to content
#

vec-perm

Here are 2 public repositories matching this topic...

Language: All
Filter by language

LLM infrastructure cost reduction via NUMA-aware weight banking: 147 t/s (8.8x stock llama.cpp) on refurbished enterprise POWER8. Self-hosted inference, no cloud APIs. Part of the Proof of Physical AI stack.

  • Updated Sep 3, 2026
  • Python

Add this topic to your repo

To associate your repository with the vec-perm topic, visit your repo's landing page and select "manage topics."

Learn more