AIPython★ 882
OpenMOSS/AnyGPT
ANYGPT — UNIFIED MULTIMODAL LANGUAGE MODEL
AnyGPT is a universal multimodal model built on discrete sequence modeling. It generates images, music, and speech from text and understands translation between modalities.
// KEY FEATURES
- Generates photorealistic images from text in seconds.
- Synthesizes natural speech and music from text prompts.
- Performs any-to-any translation: text ↔ image ↔ audio.
- Supports fine-tuning on custom multimodal datasets via a simple CLI.
#any-to-any#image-generation#large-language-models#multimodal#music-generation#speech
Open on GitHub →New repositories every 30 minutes
REDDYX AI scans GitHub 24/7 and ships the best AI/ML/Web3 projects to Telegram.
Join on Telegram