R_REDDYX.XYZ
AIPython842

mbzuai-oryx/LLaVA-pp

LLaVA++: NEW LEVEL OF MULTIMODALITY

LLaVA++ adds support for Phi‑3 and LLaMA‑3 models to the base LLaVA, enabling multimodal chat and image generation. Users gain better text‑and‑image understanding thanks to integration of cutting‑edge LLMs.

// KEY FEATURES

  • Supports multimodal chat: text + image using Phi‑3 and LLaMA‑3 models.
  • Generates answers that consider both textual and visual context at once.
  • Fine‑tunes LLaVA‑pp on custom image‑text pairs.
  • Offers a ready REST API and Python SDK for easy integration into any app.
#conversation#llama-3-llava#llama-3-vision#llama3#llama3-llava#llama3-vision#llava#llava-llama3#llava-phi3#llm#lmms#phi-3-llava
Open on GitHub →

New repositories every 30 minutes

REDDYX AI scans GitHub 24/7 and ships the best AI/ML/Web3 projects to Telegram.

Join on Telegram
← Full catalog·Full index

// SIMILAR REPOSITORIES

Claw Code: Fast Code192068n8n Process Automation187437DeepSeek Plugin Framework186631Agent AI Optimization186627AutoGPT: AI Agent184170Everything for Claude Code179304

← FULL CATALOG

mbzuai-oryx/LLaVA-pp — LLaVA++: NEW LEVEL OF MULTIMODALITY…