trending now · mentioned on @sakurayukiai
Luce-Org/lucebox
Luce-Org · Sep 5, 2026 · via trendshift · 1 min read
luce-org/lucebox
★ 2.8K⑂ 270
LLM speculative inference server for heterogeneous hardware & consumer GPUs
C++Apache-2.0#cuda#cuda-kernels#dflash#heterogeneous-computing#kernel#llama-cpp
Rendered from the publisher's own feed content.
Comments
Sign in to join the discussion.