4 comments

  • peddling-brink 1 hour ago
    > llama.cpp via Vulkan (AMD / Intel / NVIDIA) or CPU fallback

    I got excited about someone paying attention to intel. Oh well.

  • dlcarrier 59 minutes ago
    From what I've seen, Vulkan adds a lot of overhead on Intel hardware.
  • PcChip 2 hours ago
    I didn't see any benchmarks against vllm, sglang, exllama, etc
    • rancor 2 hours ago
      Since this is basically a wrapper around libllama.so, I would assume that the performance is roughly the same as llama.cpp upstream.
  • shayanjavadi 11 minutes ago
    [dead]