R_REDDYX.XYZ
AIJava276

beehive-lab/jllm

JLLM: GPU‑ACCELERATED LLAMA3.JAVA

Project jllm implements Llama3.java inference on the GPU using the TornadoVM while staying pure Java. Supports the GGUF format and models DeepSeek‑R1 and Granite. It has 276 stars on the GitHub.

// KEY FEATURES

  • Executes Llama3 model inference on GPU via TornadoVM without native code.
  • Supports loading and running models in GGUF format directly from Java.
  • Provides speedup of several times compared to pure CPU‑Java inference.
  • Compatible with DeepSeek‑R1, Granite and other LLAMA‑like checkpoints.
#accelerators#compilers#deepseek-r1#gguf#gpu#granite#ibm-granite#java#java21#llama3#llm#mistral
Open on GitHub →

New repositories every 30 minutes

REDDYX AI scans GitHub 24/7 and ships the best AI/ML/Web3 projects to Telegram.

Join on Telegram
← Full catalog·Full index

// SIMILAR REPOSITORIES

Claw Code: Fast Code192068n8n Process Automation187437DeepSeek Plugin Framework186631Agent AI Optimization186627AutoGPT: AI Agent184170Everything for Claude Code179304

← FULL CATALOG

beehive-lab/jllm — JLLM: GPU‑ACCELERATED LLAMA3.JAVA |…