\

Show HN: Janus – Go binary that runs GGUF models via Vulkan on AMD/Intel/Nvidia

50 points - yesterday at 8:36 PM

Source
  • PcChip

    yesterday at 9:19 PM

    I didn't see any benchmarks against vllm, sglang, exllama, etc

      • rancor

        yesterday at 9:27 PM

        Since this is basically a wrapper around libllama.so, I would assume that the performance is roughly the same as llama.cpp upstream.

    • dlcarrier

      yesterday at 10:40 PM

      From what I've seen, Vulkan adds a lot of overhead on Intel hardware.

      • peddling-brink

        yesterday at 10:38 PM

        > llama.cpp via Vulkan (AMD / Intel / NVIDIA) or CPU fallback

        I got excited about someone paying attention to intel. Oh well.

          • wronglebowski

            today at 1:34 AM

            What hardware do you have? I’ve been playing with a 258V and OpenVINO has come a longggggg way.

            • kamranjon

              today at 12:08 AM

              llama.cpp sycl and vllm xmx work is pretty incredible right now - you just gotta build it with some extra flags

          • shayanjavadi

            yesterday at 11:28 PM

            [dead]