XARIV Relay
Open weights on your machine — not our GPUs
Download a GGUF, deploy llama.cpp on localhost, then compare two candidates side by side. Vercel cannot host model files or llama-server, so the public site ships a timing replay for demos.
- Website (this page): split-pane speed demo, no download.
- Your Mac: real weights cache in
~/.xariv/relay/models/after the first Hugging Face fetch — Deploy does not re-download if the file is already there.