LLaVA
Multimodal modelLarge Language and Vision Assistant
Large Language and Vision Assistant is an AI model published by Microsoft Azure, released on 2023-04-17, for Multimodal model, with 13B parameters, and 2K context length, requiring about 26.1GB storage.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Model basics
Open source & experience
Official resources
API details
Benchmark Results
Compare with other models
No curated comparisons for this model yet.
Want a custom combination? Open the compare tool
Publisher
Model Overview
Large Language and Vision Assistant is an AI model published by Microsoft Azure, released on 2023-04-17, for Multimodal model, with 13B parameters, and 2K context length, requiring about 26.1GB storage.
Foundation model
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
