CLIP Interrogator
clip-interrogatorpharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
- Pris
- ≈ 0,0457 US$/kørsel
- Input → output
- Tekst + Billede → Tekst
- Udvikler
- Community
- Opdateret
- 23. september 2026
Playground
Prøv CLIP Interrogator
Input & resultat
Denne kørsel
ca. 0,0457 US$ · 4,57 credits
0,1369Â US$ (13,69 credits) reserveres ved start; den faktiske GPU-tid faktureres.
For konti uden tidligere køb: kørsler over 2 credits kræver en opladning.
Ny her?
10 gratis credits (0,10 US$) når du tilmelder dig med Google
Kan bruges 24 timer efter tilmelding, op til 5 kørsler pr. dag og højst 2 credits pr. kørsel. Andre login-metoder starter uden credits.
Examples
InputNo text prompt: the model only takes the input shown.
Response
a watercolor painting of a sea turtle, a digital painting, by Kubisi art, featured on dribbble, medibang, warm saturated palette, red and green tones, turquoise horizon, digital art h 9 6 0, detailed scenery —width 672, illustration:.4, spray art, artstatiom
InputNo text prompt: the model only takes the input shown.
Response
a painting of a field with green grass, by Vincent Van Gogh, featured on deviantart, windy day, virtuosic level detail, the front of a trading card, loosely cropped, expansive, tendrils in the background, photo courtesy museum of art, 1980s art, standing on a hill, creative commons attribution
Om CLIP Interrogator
CLIP Interrogator er en model af Community i kategorien Multimodal. På Railwail koster CLIP Interrogator ≈ 0,0457 US$ pr. kørsel.
Priser
| Typisk kørsel (≈ 169 s på T4) | 0,0457 US$ pr. kørsel |
|---|---|
| GPU-tid (T4) | 0,00027Â US$ pr. GPU-sekund |
- Faktureres efter den GPU-tid, som kørslen faktisk tager. Når kørslen starter, reserveres 3× den typiske pris fra din saldo og afregnes efterfølgende.
- 1 kredit = 0,01Â US$
Omkostningsberegner
Prisberegner
Typisk ifølge udbyderen: ca. 168,9 s
I alt
4,57Â US$
457 credits
Pr. kørsel
0,0457 US$ · 4,57 credits
Fakturering efter faktisk GPU-tid; dette er et estimat.
API
Intet bekræftet API-eksempel
Den offentlige API sender et andet inputformat end det, som denne model har brug for. Brug playground ovenfor.
Specifikationer
- Model-ID
clip-interrogator- Udvikler
- Community
- Kategori
- Multimodal
- Input
- Tekst, Billede
- Output
- Tekst
- Fakturering
- Efter forbrug (tokens eller GPU-tid)
- Katalogelement opdateret
- 23. september 2026
Inputparametre
Inputs og indstillinger fra modellens inputskema. Eksemplet i API-afsnittet viser, hvilke af dem API'en accepterer.
imagepåkrævetImage to interrogate
Type: TekstStandard: –Tilladte værdier: –modeType: ValgStandard:bestTilladte værdier: best, fast, classic eller negativeclip_model_nameType: TekstStandard:ViT-L-14/openaiTilladte værdier: –
Tags
- replicate
- clip-interrogator
- captioning
- tagging
- clip
- blip
- prompt-generation
- image
Ofte stillede spørgsmål
Hvad er CLIP Interrogator?
CLIP Interrogator er en model fra Community i kategorien Multimodal.
Hvad koster CLIP Interrogator på Railwail?
På Railwail koster CLIP Interrogator ≈ 0,0457 US$ pr. kørsel. Du betaler for det, som hver anmodning faktisk bruger. Forbrug betales fra forudbetalte credits; 1 credit svarer til 0,01 US$.
Hvilke indstillinger understøtter CLIP Interrogator?
Ifølge dens inputskema kender CLIP Interrogator disse parametre: image, mode (best, fast, classic eller negative) og clip_model_name.
Hvor hurtig er CLIP Interrogator?
Der er endnu ikke nok målte kørsler af CLIP Interrogator på Railwail til at angive en udførelsestid. Det afhænger af inputtet, indstillingerne og belastningen hos provideren.
Er CLIP Interrogator bedre end BLIP?
Det afhænger af opgaven. CLIP Interrogator (Community) og BLIP (Salesforce) er begge modeller i kategorien Multimodal. Sammenligningssiden viser deres priser og specifikationer side om side.
Sammenlign CLIP Interrogator og BLIPKan CLIP Interrogator behandle billeder?
Ja. CLIP Interrogator accepterer billeder som input ud over tekst.
Sammenlignelige modeller
Alle i denne kategori- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- Depth Anything v2Community
Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
Meta Segment Anything 2. Promptable segmentation across images and video with temporal memory. Zero-shot, point/box/mask prompts, fast on a single H100.
≈ 0,018 US$/kørsel
61 % billigere pr. enhed
Sammenlign CLIP Interrogator og SAM 2 (Segment Anything 2)
Alle modeller via én API
En API-nøgle til alle modeller på Railwail. Forbrug debiteres fra forudbetalte credits, 1 credit = 0,01 US$.