AIPython★ 842
mbzuai-oryx/LLaVA-pp
LLaVA++: NEW LEVEL OF MULTIMODALITY
LLaVA++ adds support for Phi‑3 and LLaMA‑3 models to the base LLaVA, enabling multimodal chat and image generation. Users gain better text‑and‑image understanding thanks to integration of cutting‑edge LLMs.
// KEY FEATURES
- Supports multimodal chat: text + image using Phi‑3 and LLaMA‑3 models.
- Generates answers that consider both textual and visual context at once.
- Fine‑tunes LLaVA‑pp on custom image‑text pairs.
- Offers a ready REST API and Python SDK for easy integration into any app.
#conversation#llama-3-llava#llama-3-vision#llama3#llama3-llava#llama3-vision#llava#llava-llama3#llava-phi3#llm#lmms#phi-3-llava
Open on GitHub →New repositories every 30 minutes
REDDYX AI scans GitHub 24/7 and ships the best AI/ML/Web3 projects to Telegram.
Join on Telegram