Moondream2
moondream2Moondream2 small vision-language model on Replicate. About 1.9B params, designed to run on edge devices, handles captioning, visual QA and short OCR-style reads at very low cost.
- Preis
- ca. 0,0020 $/Lauf
- Eingabe → Ausgabe
- Text + Bild → Text
- Entwickler
- Community
- Aktualisiert
- 23. September 2026
Playground
Moondream2 ausprobieren
Eingabe & Ergebnis
Dieser Lauf
ca. 0,002 $ · 0,2 Credits
Beim Start werden 0,0058 $ (0,58 Credits) vorgemerkt, abgerechnet wird die tatsächliche GPU-Zeit.
Neu hier?
10 Gratis-Credits (0,10 $) bei Anmeldung mit Google
Nutzbar 24 Stunden nach der Anmeldung, bis zu 5 Läufe pro Tag und höchstens 2 Credits je Lauf. Andere Anmeldearten starten ohne Guthaben. Reicht für 17 Läufe dieses Modells.
Beispiele
EingabePrompt
Describe this image
Antwort
The image features a logo with a smiling blue circle above the word "moondream" written in black text.
EingabePrompt
Describe this image
Antwort
A man with a beard and mustache, wearing a suit and red tie, is smiling at the camera with a blue background featuring a logo.
EingabePrompt
Describe this image
Antwort
An astronaut in a white spacesuit is riding a unicorn with a rainbow mane and tail, soaring through a colorful, dreamlike sky with clouds and rainbows.
Über Moondream2
Moondream2 ist ein Modell von Community aus der Kategorie Multimodal. Über Railwail kostet Moondream2 ca. 0,0020 $ pro Lauf.
Preise
| Typischer Lauf (ca. 2 s auf L40S) | 0,0020 $ pro Lauf |
|---|---|
| GPU-Zeit (L40S) | 0,00117 $ pro GPU-Sekunde |
- Abgerechnet wird die GPU-Zeit, die der Lauf tatsächlich braucht. Beim Start wird das 3-Fache des typischen Preises vom Guthaben vorgemerkt und danach verrechnet.
- 1 Credit = 0,01 $
Kostenrechner
Preisrechner
Typisch laut Anbieter: ca. 1,6 s
Gesamt
0,20 $
20 Credits
Je Lauf
0,002 $ · 0,2 Credits
Abgerechnet wird die tatsächliche GPU-Zeit; der Wert ist eine Schätzung.
API
Kein geprüftes API-Beispiel
Die öffentliche API übergibt ein anderes Eingabeformat, als dieses Modell braucht. Nutze den Playground oben.
Spezifikationen
- Modell-ID
moondream2- Entwickler
- Community
- Kategorie
- Multimodal
- Eingabe
- Text, Bild
- Ausgabe
- Text
- Abrechnung
- Nach Verbrauch (Token bzw. GPU-Zeit)
- Katalogeintrag aktualisiert
- 23. September 2026
Eingabeparameter
Eingaben und Einstellungen laut Eingabeschema des Modells. Welche davon die API annimmt, zeigt das Beispiel im Abschnitt API.
promptPflichtQuestion about the image
Typ: TextStandard: –Erlaubte Werte: bis 8.000 Zeichenimage_urlImage URL to analyze
Typ: TextStandard: –Erlaubte Werte: –
Schlagwörter
- replicate
- moondream
- vision-understanding
- open-source
- small
- edge
Häufige Fragen
Was ist Moondream2?
Moondream2 ist ein Modell von Community aus der Kategorie Multimodal.
Was kostet Moondream2 bei Railwail?
Über Railwail kostet Moondream2 ca. 0,0020 $ pro Lauf. Abgerechnet wird, was jede Anfrage tatsächlich verbraucht. Bezahlt wird mit vorab gekauften Credits; 1 Credit entspricht 0,01 $.
Welche Einstellungen unterstützt Moondream2?
Laut Eingabeschema kennt Moondream2 diese Parameter: prompt (bis 8.000 Zeichen) und image_url.
Wie schnell ist Moondream2?
Für Moondream2 gibt es bei Railwail noch zu wenige gemessene Läufe, um eine Laufzeit anzugeben. Sie hängt von der Eingabe, den Einstellungen und der Auslastung beim Anbieter ab.
Ist Moondream2 besser als BLIP?
Das hängt von der Aufgabe ab. Moondream2 (Community) und BLIP (Salesforce) sind beide Modelle aus der Kategorie Multimodal. Die Vergleichsseite zeigt Preise und Spezifikationen nebeneinander.
Moondream2 und BLIP vergleichenKann Moondream2 Bilder verarbeiten?
Ja. Moondream2 nimmt neben Text auch Bilder als Eingabe an.
Vergleichbare Modelle
Alle dieser Kategorie- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- CLIP InterrogatorCommunity
pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
- Depth Anything v2Community
Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
Alle Modelle über eine API
Ein API-Schlüssel für alle Modelle auf Railwail. Abgerechnet wird über vorab gekaufte Credits, 1 Credit = 0,01 $.