Llama 3.2 Vision 90B
llama-3-2-vision-90bMeta Llama 3.2 90B Vision. Largest open-weights Llama vision model. Strong visual reasoning, chart, OCR and document understanding.
- Prix
- ≈ 0,0070 $US/exécution
- Contexte
- 131 072 tokens
- Max. sortie
- 4 096 tokens
- Entrée → Sortie
- Texte + Image → Texte
- Développeur
- Community
- Mis à jour
- 23 septembre 2026
Playground
Essayer Llama 3.2 Vision 90B
Pas de formulaire d'entrée
Aucun formulaire d'entrée pour ce modèle pour le moment
Ses entrées ne sont pas encore documentées. Pour éviter qu'une exécution échoue avec une mauvaise entrée, nous ne proposons pas de formulaire ici. Choisissez plutôt un modèle comparable.
Examples
InputPrompt
Which era does this piece belong to? Give details about the era.
Response
The image features a ceiling painting or mural, which is characteristic of the Rococo style. This art movement emerged in Europe during the 18th century, roughly between 1720 and 1770. **Key Characteristics:** * Lighter, more delicate colors compared to Baroque art * Use of pastel shades and soft hues **Associated Periods:** * Late Baroque period (1685-1750) * Early Classical period (1750-1820) These periods are marked by a shift away from the grandeur and drama of Baroque art towards a more refined and elegant aesthetic.
InputPrompt
Where to find this monument? Can you give me other recommendations around the area?
Response
The monument in the image is located at Wat Phra Kaew, also known as the Temple of the Emerald Buddha, which is a complex of buildings within the Grand Palace in Bangkok, Thailand. The temple is situated on Na Phra Lan Road in Phra Nakhon District, Bangkok. **Other Recommendations Around the Area:** * **The Grand Palace**: A must-visit attraction in Bangkok, the palace was the official residence of the king and queen of Thailand until 1925. * **Wat Arun (Temple of Dawn)**: Located on the west bank of the Chao Phraya River, this temple is known for its stunning architecture and beautiful views of the river. * **Wat Pho (Temple of the Reclining Buddha)**: Home to a massive reclining Buddha statue, Wat Pho is one of the largest temples in Bangkok and offers traditional Thai massage services. * **Boat Ride along the Chao Phraya River**: Take a boat ride along the river to see the city from a different perspective and visit other attractions such as Wat Arun and Wat Pho. * **Floating Markets**: Visit one of the many floating markets in Bangkok, such as Damnoen Saduak Floating Market or Taling Chan Floating Market, to experience traditional Thai culture and try local food. * **Street Food**: Try some of the delicious street food available throughout Bangkok, including dishes like pad thai, tom yum soup, and mango sticky rice. * **MBK Center**: A large shopping mall with over 2,000 stores selling everything from clothing to electronics to souvenirs. * **Siam Paragon**: Another po…
InputPrompt
What happens when you take out white cat from this chain?
Response
**Food Web Analysis: Impact of Removing the Wild Cat** To understand the effects of removing the wild cat from this food web, we need to examine its role and connections within the ecosystem. **Position in the Food Web** * The wild cat is a predator that feeds on rabbits. * Rabbits are preyed upon by owls but also consume green plants. **Impact on Prey Population (Rabbits)** * Removing the wild cat would reduce predation pressure on rabbits. * This could lead to an increase in the rabbit population, as one of their predators is removed. **Potential Effects on Other Species** * An increased rabbit population could result in: + Overgrazing: More rabbits consuming green plants could lead to a decrease in plant biomass and diversity. + Impact on Owl Population: With fewer wild cats competing for rabbit prey, the owl population might increase due to the abundance of rabbits. **Conclusion** Removing the wild cat from this food web would likely cause an increase in the rabbit population. This, in turn, could lead to overgrazing by rabbits and potentially impact the owl population positively due to increased availability of their primary food source.
À propos de Llama 3.2 Vision 90B
Llama 3.2 Vision 90B est un modèle de Community dans la catégorie Multimodal. Sur Railwail, Llama 3.2 Vision 90B coûte ≈ 0,0070 $US par exécution. La fenêtre de contexte contient 131 072 tokens, et une réponse peut faire jusqu'à 4 096 tokens.
Tarification
| Exécution typique (≈ 4 s sur A100 (80GB)) | 0,0070 $US par exécution |
|---|---|
| Temps GPU (A100 (80GB)) | 0,00168Â $US par seconde GPU |
- Facturée selon le temps GPU que l'exécution prend réellement. Au démarrage, 3× le prix typique est réservé de votre solde et régularisé ensuite.
- 1 crédit = 0,01 $US
Calculatrice de coûts
Calculatrice de prix
Typique selon le fournisseur : environ 4,1 s
Total
0,70Â $US
70 crédits
Par exécution
0,007 $US · 0,7 crédits
Facturé selon le temps GPU réel ; ceci est une estimation.
API
Aucun exemple API vérifié
L'API publique transmet un format d'entrée différent de celui dont ce modèle a besoin. Utilisez le playground ci-dessus.
Spécifications
- ID du modèle
llama-3-2-vision-90b- Développeur
- Community
- Catégorie
- Multimodal
- Entrée
- Texte, Image
- Sortie
- Texte
- Fenêtre de contexte
- 131 072 tokens
- Sortie max.
- 4 096 tokens
- Facturation
- À l'usage (tokens ou temps GPU)
- Entrée du catalogue mise à jour
- 23 septembre 2026
Étiquettes
- replicate
- multimodal
- vision-understanding
- meta
- open-weights
Questions fréquemment posées
Qu'est-ce que Llama 3.2 Vision 90B ?
Llama 3.2 Vision 90B est un modèle de Community dans la catégorie Multimodal.
Combien coûte Llama 3.2 Vision 90B sur Railwail ?
Sur Railwail, Llama 3.2 Vision 90B coûte ≈ 0,0070 $US par exécution. Vous êtes facturé pour ce que chaque requête utilise réellement. L'utilisation est payée à partir de crédits prépayés ; 1 crédit équivaut à 0,01 $US.
Quelle est la fenêtre de contexte de Llama 3.2 Vision 90B ?
La fenêtre de contexte de Llama 3.2 Vision 90B contient 131 072 tokens. Une réponse peut faire jusqu'à 4 096 tokens.
Quelle est la vitesse de Llama 3.2 Vision 90B ?
Il n'y a pas encore assez d'exécutions mesurées de Llama 3.2 Vision 90B sur Railwail pour indiquer un temps d'exécution. Cela dépend de l'entrée, des paramètres et de la charge chez le fournisseur.
Llama 3.2 Vision 90B est-il meilleur que BLIP ?
Cela dépend de la tâche. Llama 3.2 Vision 90B (Community) et BLIP (Salesforce) sont tous deux des modèles de la catégorie Multimodal. La page de comparaison affiche leurs prix et spécifications côte à côte.
Comparer Llama 3.2 Vision 90B et BLIPLlama 3.2 Vision 90B peut-il traiter des images ?
Oui. Llama 3.2 Vision 90B accepte les images en entrée en plus du texte.
Modèles comparables
Tous dans cette catégorie- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- CLIP InterrogatorCommunity
pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
≈ 0,0457 $US/exécution
553 % plus cher par unité
Comparer Llama 3.2 Vision 90B et CLIP Interrogator - Depth Anything v2Community
Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
≈ 0,0050 $US/exécution
29 % moins cher par unité
Comparer Llama 3.2 Vision 90B et Depth Anything v2
Tous les modèles via une API
Une clé API pour tous les modèles sur Railwail. L'utilisation est facturée à partir de crédits prépayés, 1 crédit = 0,01 $US.