N
Hacker Next
new
past
show
ask
show
jobs
submit
login
▲
Show HN: Janus – Go binary that runs GGUF models via Vulkan on AMD/Intel/Nvidia
(
github.com
)
61 points by
Maverick617
9 hours ago
|
9 comments
add comment
Rendered at 05:35:13 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
PcChip 8 hours ago
[-]
I didn't see any benchmarks against vllm, sglang, exllama, etc
rancor 8 hours ago
[-]
Since this is basically a wrapper around libllama.so, I would assume that the performance is roughly the same as llama.cpp upstream.
nullpoint420 1 hours ago
[-]
Woof. Wonder if the creator knows that
dlcarrier 7 hours ago
[-]
From what I've seen, Vulkan adds a lot of overhead on Intel hardware.
peddling-brink 7 hours ago
[-]
> llama.cpp via Vulkan (AMD / Intel / NVIDIA) or CPU fallback
I got excited about someone paying attention to intel. Oh well.
wronglebowski 4 hours ago
[-]
What hardware do you have? I’ve been playing with a 258V and OpenVINO has come a longggggg way.
peddling-brink 3 hours ago
[-]
Two arc b60s. The intel vllm build is getting me ~15t/s decode with heavy context using qwen3.8 27b.
kamranjon 5 hours ago
[-]
llama.cpp sycl and vllm xmx work is pretty incredible right now - you just gotta build it with some extra flags
peddling-brink 3 hours ago
[-]
Llama would be nice for the ggufs. Any specific flags or tutorials I should look at?
shayanjavadi 6 hours ago
[-]
[dead]
I got excited about someone paying attention to intel. Oh well.