Om denna modell
TEXT / LLM
meta-llama | 32,768 token context | $0.35/M input, $0.40/M output. The Llama 90B Vision model is a top-tier, 90-billion-parameter multimodal model designed for the most challenging visual reasoning and language tasks. It offers unparalleled accuracy in image captioning, visual question answering, and advanced image-text comprehension. Pre-trained on vast multimodal datasets and fine-tuned with human feedback, the Llama 90B Vision is engineered to handle the most demanding image-based AI tasks. This model is perfect for industries requiring cutting-edge multimodal AI capabilities, particularly those dealing with complex, real-time visual and textual analysis.
ModelTrends.ai modell-ID
#27984
Rapportera innehåll
AI-modeller för bild-, video- och textgenerering
Llama 3.2 90B Vision Instruct är en textmodell för familjen Text / LLM, listad på ModelTrends.ai, en skrivskyddad katalog över AI-modeller med öppen källkod för bild, video och text. Jämför den med andra Text / LLM-modeller, se dess hettapoäng och prova den på PirateDiffusion eller BitVector.
Bläddra efter modellfamilj
Text / LLM-modeller
411
Anima-modeller
509
Chroma-modeller
14
Flux-modeller
1,071
Flux 2 / Klein-modeller
1,237
MiniMax H3-modeller
331
Hunyuan-modeller
197
Ideogram-modeller
20
Krea2-modeller
674
Qwen-modeller
17
Qwen2-modeller
36
Ltx2-modeller
150
Stable Diffusion 1.5-modeller
6,465
Stable Diffusion XL 1.0-modeller
15,432
Zimage-modeller
804
Wan-modeller
919

Model
Trends
.ai
© ModelTrends.ai
|
Tillverkad i Japan
|
© 2026
