AIJava★ 276
beehive-lab/jllm
JLLM: GPU‑ACCELERATED LLAMA3.JAVA
Project jllm implements Llama3.java inference on the GPU using the TornadoVM while staying pure Java. Supports the GGUF format and models DeepSeek‑R1 and Granite. It has 276 stars on the GitHub.
// KEY FEATURES
- Executes Llama3 model inference on GPU via TornadoVM without native code.
- Supports loading and running models in GGUF format directly from Java.
- Provides speedup of several times compared to pure CPU‑Java inference.
- Compatible with DeepSeek‑R1, Granite and other LLAMA‑like checkpoints.
#accelerators#compilers#deepseek-r1#gguf#gpu#granite#ibm-granite#java#java21#llama3#llm#mistral
Open on GitHub →New repositories every 30 minutes
REDDYX AI scans GitHub 24/7 and ships the best AI/ML/Web3 projects to Telegram.
Join on Telegram