LLaVA 1.6 Vicuna 13B
llava-1-6-vicuna-13bLLaVA 1.6 (LLaVA-NeXT) with a Vicuna-13B language backbone. Open vision-language chat model that describes images, answers questions, reads charts and reasons about scenes. Version 1.6 adds higher input resolution and better OCR and reasoning than LLaVA 1.5.
- Preț
- ≈ 0,1021 USD/rulare
- Context
- 4.096 tokeni
- Ieșire max.
- 1.024 tokeni
- Intrare → ieșire
- Text + Imagine → Text
- Dezvoltator
- Community
- Actualizat
- 23 septembrie 2026
Playground
Încearcă LLaVA 1.6 Vicuna 13B
Intrare & rezultat
Această rulare
aproximativ 0,1021 USD · 10,21 credite
La start se rezervă 0,3061 USD (30,61 credite); se factureaza timpul GPU real.
Pentru conturi fără achiziție anterioară: rulările peste 2 credite necesită o reîncărcare.
Nou aici?
10 credite gratuite (0,10 USD) când te înregistrezi cu Google
Utilizabil 24 ore după înregistrare, până la 5 rulări pe zi și maximum 2 credite pe rulare. Alte metode de conectare încep fără credite.
Examples
InputPrompt
What should I take into account when visiting this place?
Response
When visiting a place like the one shown in the image, which appears to be a serene lake with a dock and surrounded by mountains and forests, you should consider the following: 1. **Weather Conditions**: Be prepared for the local weather. It can change rapidly, especially in mountainous areas, so check the forecast and dress in layers. 2. **Safety**: If you plan to go on the water, make sure you have a life jacket and know how to swim. If you're not a strong swimmer, consider wearing a life jacket even if you're just walking along the shore. 3. **Wildlife**: Be aware of the local wildlife and maintain a safe distance. Do not feed the animals, as this can disrupt their natural behavior and diet. 4. **Leave No Trace**: Respect the environment by not littering and by staying on designated paths or trails. 5. **Emergency Preparedness**: Have a first aid kit and know basic first aid. It's also a good idea to have a way to contact emergency services if needed. 6. **Navigation**: If you're going into the woods or on the water, have a map or GPS device and know how to use it. 7. **Hydration and Nutrition**: Bring enough water and food for your visit, especially if you'll be active. 8. **Clothing and Gear**: Wear appropriate clothing for the weather and activities you plan to do. Bring sturdy footwear if you plan to hike or walk on uneven terrain. 9. **Local Regulations**: Follow any posted rules or regulations, which might include fishing regulations, fire restrictions, or…
InputPrompt
What is unusual about this image?
Response
The image shows a person ironing clothes on the back of a moving vehicle, which is an unusual and potentially dangerous activity. Ironing clothes while a vehicle is in motion can be hazardous for the person doing the ironing as well as for other road users. The person is at risk of losing balance and falling off the vehicle, which could result in serious injury. Additionally, the iron could potentially cause a fire or damage to the vehicle or its surroundings if it overheats or malfunctions. This scene is not a safe or typical way to iron clothes and is likely staged for comedic or dramatic effect.
Despre LLaVA 1.6 Vicuna 13B
LLaVA 1.6 Vicuna 13B este un model de Community din categoria Multimodal. Pe Railwail, LLaVA 1.6 Vicuna 13B costă ≈ 0,1021 USD per rulare. Fereastra de context conține 4.096 token-uri, iar un răspuns poate fi lung de până la 1.024 token-uri.
Prețuri
| Rulare tipică (≈ 87 s pe L40S) | 0,1021 USD per rulare |
|---|---|
| Timp GPU (L40S) | 0,00117 USD per secundă GPU |
- Se facturează timpul GPU pe care îl necesită efectiv rularea. La pornire, se rezervă din soldul tău 3× prețul tipic și se decontează ulterior.
- 1 credit = 0,01 USD
Calculator de costuri
Calculator de preț
Tipic conform furnizorului: aprox. 87,2 s
Total
10,21 USD
1.021 credite
Pe rulare
0,1021 USD · 10,21 credite
Se factură după timpul GPU real; aceasta este o estimare.
API
Niciun exemplu API verificat
API-ul public transmite un format de intrare diferit de ceea ce are nevoie acest model. Folosește playground-ul de mai sus.
Specificații
- ID model
llava-1-6-vicuna-13b- Dezvoltator
- Community
- Categorie
- Multimodal
- Intrare
- Text, Imagine
- Ieșire
- Text
- Fereastră de context
- 4.096 tokeni
- Ieșire max.
- 1.024 tokeni
- Facturare
- După utilizare (tokeni sau timp GPU)
- Intrare catalog actualizată
- 23 septembrie 2026
Parametri de intrare
Intrări și setări din schema de intrare a modelului. Exemplul din secțiunea API arată care dintre ele acceptă API-ul.
imageobligatoriuImage to analyze
Tip: TextImplicit: –Valori permise: –promptobligatoriuQuestion or instruction about the image
Tip: TextImplicit:Describe this image in detail.Valori permise: până la 4.000 caracteretop_pTip: NumărImplicit:1Valori permise: 0 până la 1max_tokensTip: Număr întregImplicit:512Valori permise: 1 până la 1.024temperatureTip: NumărImplicit:0.2Valori permise: 0 până la 2
Etichete
- replicate
- llava
- captioning
- vqa
- vision-understanding
- open-weights
- image
Întrebări frecvente
Ce este LLaVA 1.6 Vicuna 13B?
LLaVA 1.6 Vicuna 13B este un model de Community din categoria Multimodal.
Cât costă LLaVA 1.6 Vicuna 13B pe Railwail?
Pe Railwail, LLaVA 1.6 Vicuna 13B costă ≈ 0,1021 USD per rulare. Ți se percepe taxa pentru ceea ce fiecare cerere folosește efectiv. Utilizarea se plătește din credite prepay; 1 credit egal cu 0,01 USD.
Care este fereastra de context a LLaVA 1.6 Vicuna 13B?
Fereastra de context a LLaVA 1.6 Vicuna 13B conține 4.096 token-uri. Un răspuns poate fi lung de până la 1.024 token-uri.
Cât de rapid este LLaVA 1.6 Vicuna 13B?
Nu sunt suficiente rulări măsurate ale LLaVA 1.6 Vicuna 13B pe Railwail încă pentru a indica un timp de rulare. Depinde de intrare, de setări și de sarcina la furnizor.
Este LLaVA 1.6 Vicuna 13B mai bun decât BLIP?
Depinde de sarcină. LLaVA 1.6 Vicuna 13B (Community) și BLIP (Salesforce) sunt ambele modele din categoria Multimodal. Pagina de comparație arată prețurile și specificațiile lor una lângă alta.
Compară LLaVA 1.6 Vicuna 13B și BLIPPoate LLaVA 1.6 Vicuna 13B procesa imagini?
Da. LLaVA 1.6 Vicuna 13B acceptă imagini ca intrare, pe lângă text.
Modele comparabile
Toate din această categorie- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- CLIP InterrogatorCommunity
pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
- Depth Anything v2Community
Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
Toate modelele printr-o singură API
O cheie API pentru fiecare model pe Railwail. Utilizarea se percepe din credite prepay, 1 credit = 0,01 USD.