Llama 3.2 Vision 11B (Ollama)
llama-3-2-vision-11b-ollamaMeta Llama 3.2 11B Vision served via Ollama on Replicate. Open-weights multimodal model for image captioning, document and chart reading, and visual question answering.
- Ціна
- ≈ 0,0039 USD/запуск
- Вхід → вихід
- Текст + Зображення → Текст
- Розробник
- Community
- Оновлено
- 23 вересня 2026 р.
Playground
Спробувати Llama 3.2 Vision 11B (Ollama)
Введення та результат
Цей запуск
близько 0,0039 USD · 0,39 кредитів
На початку зарезервовано 0,0116 USD (1,16 кредитів); розраховується фактичний час GPU.
Новенький?
10 безплатних кредитів (0,10 USD) при реєстрації через Google
Доступно 24 годин після реєстрації, до 5 запусків на день і максимум 2 кредитів за запуск. Інші способи входу стартують без кредитів. Достатньо для 8 запусків цієї моделі.
Examples
InputPrompt
Where to find this monument? Can you give me other recommendations around the area?
Response
**Wat Phra Kaew (Temple of the Emerald Buddha)** in Bangkok, Thailand. Located within the Grand Palace complex, Wat Phra Kaew is a highly revered and sacred site that houses the iconic Emerald Buddha image. The temple's design incorporates traditional Thai architecture with intricate details and ornate decorations, featuring an imposing spire topped by a golden finial adorned with images of mythical creatures. **Recommendations for Your Visit:** * **Grand Palace**: Explore the sprawling palace grounds, which include several other notable temples and buildings. * **Wat Arun (Temple of Dawn)**: A nearby temple famous for its distinctive spire covered in colorful ceramics and glints of gold. * **Chao Phraya River Cruise**: Take a boat ride along the river to admire the city's skyline and learn about its history. * **Jim Thompson House Museum**: Visit this museum dedicated to Thai art and culture, featuring traditional architecture and beautiful gardens. Wat Phra Kaew is an absolute must-visit attraction when in Bangkok. With these recommendations, you'll have a well-rounded experience of the city's rich cultural heritage.
InputPrompt
How long does it take from invoice date to due date? Be short and concise.
Response
From the invoice date of 11/02/2019 to the due date of 26/02/2019, there is a difference of **15 days**. This calculation was performed by subtracting the invoice date from the due date: 26/02/2019 - 11/02/2019 = 15
InputPrompt
What happens when you take out white cat from this chain?
Response
The white cat is not present in the food web. The correct term for a wild animal would be "wildcat." If we were to remove a wildcat from the ecosystem, it would no longer be able to hunt or prey on other animals. This could have cascading effects throughout the entire ecosystem. *Answer*: Wildcat will lose its source of food and nutrients
Про Llama 3.2 Vision 11B (Ollama)
Llama 3.2 Vision 11B (Ollama) — це модель від Community у категорії Мультимодальні. На Railwail Llama 3.2 Vision 11B (Ollama) коштує ≈ 0,0039 USD за запуск.
Ціни
| Типовий запуск (≈ 3 с на L40S) | 0,0039 USD за запуск |
|---|---|
| GPU-час (L40S) | 0,00117 USD за GPU-секунду |
- Оплачується за GPU-час, який насправді займає запуск. При запуску 3× типової ціни зарезервовується з вашого балансу та розраховується потім.
- 1 кредит = 0,01 USD
Калькулятор вартості
Калькулятор ціни
Типово за постачальником: близько 3,3 с
Всього
0,39 USD
39 кредитів
За запуск
0,0039 USD · 0,39 кредитів
Виставляється рахунок за фактичний час GPU; це оцінка.
API
Немає перевіреного прикладу API
Публічний API передає інший формат введення, ніж потребує ця модель. Використовуйте playground вище.
Характеристики
- ID моделі
llama-3-2-vision-11b-ollama- Розробник
- Community
- Категорія
- Мультимодальні
- Вхідні дані
- Текст, Зображення
- Вихідні дані
- Текст
- Виставлення рахунків
- За використанням (токени або час GPU)
- Запис у каталозі оновлено
- 23 вересня 2026 р.
Вхідні параметри
Вхідні дані та налаштування зі схеми вхідних даних моделі. Приклад у розділі API показує, які з них приймає API.
promptобов'язковоQuestion about the image
Тип: ТекстЗа замовчуванням: –Допустимі значення: до 16 000 символівimage_urlImage URL to analyze
Тип: ТекстЗа замовчуванням: –Допустимі значення: –max_tokensТип: Ціле числоЗа замовчуванням:1024Допустимі значення: 1 до 4 096temperatureТип: ЧислоЗа замовчуванням:0.7Допустимі значення: 0 до 2
Теги
- replicate
- meta
- llama
- vision-understanding
- open-weights
- ollama
Часті питання
Що таке Llama 3.2 Vision 11B (Ollama)?
Llama 3.2 Vision 11B (Ollama) — це модель від Community у категорії Мультимодальні.
Скільки коштує Llama 3.2 Vision 11B (Ollama) на Railwail?
На Railwail Llama 3.2 Vision 11B (Ollama) коштує ≈ 0,0039 USD за запуск. Вам буде виставлено рахунок за те, що насправді використовує кожен запит. Використання оплачується з попередньо придбаних кредитів; 1 кредит дорівнює 0,01 USD.
Які налаштування підтримує Llama 3.2 Vision 11B (Ollama)?
Згідно зі схемою введення, Llama 3.2 Vision 11B (Ollama) знає такі параметри: prompt (до 16 000 символів), image_url, max_tokens (1 до 4 096) і temperature (0 до 2).
Наскільки швидкий Llama 3.2 Vision 11B (Ollama)?
На Railwail поки що недостатньо виміряних запусків Llama 3.2 Vision 11B (Ollama), щоб вказати час виконання. Це залежить від введення, налаштувань та навантаження у постачальника.
Чи Llama 3.2 Vision 11B (Ollama) краще за BLIP?
Це залежить від завдання. Llama 3.2 Vision 11B (Ollama) (Community) і BLIP (Salesforce) — обидві моделі в категорії Мультимодальні. На сторінці порівняння показані їхні ціни та специфікації поруч.
Порівняти Llama 3.2 Vision 11B (Ollama) і BLIPЧи може Llama 3.2 Vision 11B (Ollama) обробляти зображення?
Так. Llama 3.2 Vision 11B (Ollama) приймає зображення як введення, крім тексту.
Порівнювані моделі
Усі в цій категорії- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- CLIP InterrogatorCommunity
pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
≈ 0,0457 USD/запуск
1 072 % дорожче за одиницю
Порівняти Llama 3.2 Vision 11B (Ollama) та CLIP Interrogator - Depth Anything v2Community
Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
≈ 0,0050 USD/запуск
28 % дорожче за одиницю
Порівняти Llama 3.2 Vision 11B (Ollama) та Depth Anything v2
Усі моделі через один API
Один API-ключ для всіх моделей на Railwail. Використання оплачується з попередньо поповнених кредитів, 1 кредит = 0,01 USD.