Settings

Theme

Show HN: Janus – Go binary that runs GGUF models via Vulkan on AMD/Intel/Nvidia

github.com

44 points by Maverick617 · 6 comments

Reader

3 threads
dlcarrier

From what I've seen, Vulkan adds a lot of overhead on Intel hardware.

PcChip

I didn't see any benchmarks against vllm, sglang, exllama, etc

  • rancor

    Since this is basically a wrapper around libllama.so, I would assume that the performance is roughly the same as llama.cpp upstream.

peddling-brink

> llama.cpp via Vulkan (AMD / Intel / NVIDIA) or CPU fallback

I got excited about someone paying attention to intel. Oh well.

  • kamranjon

    llama.cpp sycl and vllm xmx work is pretty incredible right now - you just gotta build it with some extra flags

Keyboard Shortcuts

j
Next item
k
Previous item
o / Enter
Open selected item
?
Show this help
Esc
Close modal / clear selection