Dots OCR

MultimodaleDisponibile
di CommunityID modello: dots-ocr

Rednote Hilab Dots OCR. End-to-end document parsing model with layout, text and reading-order prediction in one transformer.

Prezzo
≈ 0,0132 USD/esecuzione
Input → output
Testo + Immagine → Testo
Sviluppatore
Community
Aggiornato
23 settembre 2026
01

Playground

Prova Dots OCR

Nessuna maschera di input

≈ 0,0132 USD/esecuzione

Nessun modulo di input per questo modello

I suoi input non sono ancora documentati. Per evitare errori di esecuzione, non offriamo un modulo qui. Scegli invece un modello comparabile.

02

Examples

Real outputs from the public examples of this model on Replicate, with the prompt and settings that produced them. They were not generated live on this page.
  • InputInput

    Prompt

    grounding_ocr

    Response

    [{"bbox": [448, 513, 528, 535], "category": "Section-header", "text": "### Directions"}, {"bbox": [448, 533, 824, 565], "category": "Text", "text": "For sunscreen use: • apply generously 15 minutes before sun exposure • reapply: • after 80 minutes of swimming or sweating • immediately after towel drying • at least every 2 hours"}, {"bbox": [448, 563, 832, 624], "category": "List-item", "text": "• **Sun Protection Measures.** Spending time in the sun increases your risk of skin cancer and early skin aging. To decrease this risk, regularly use a sunscreen with a Broad Spectrum SPF value of 15 or higher and other sun protection measures including: • limit time in the sun, especially from 10 a.m. - 2 p.m. • wear long-sleeved shirts, pants, hats and sunglasses"}, {"bbox": [448, 622, 640, 638], "category": "List-item", "text": "• children under 6 months of age: Ask a doctor"}]

03

Informazioni su Dots OCR

RiassuntoA partire da 23 settembre 2026

Dots OCR è un modello di Community nella categoria Multimodale. Su Railwail, Dots OCR costa ≈ 0,0132 USD per esecuzione.

04

Prezzi

Prezzi in dollari USA. L'utilizzo viene addebitato dai crediti prepagati.
Esecuzione tipica (≈ 11 s su L40S)0,0132 USD per esecuzione
Tempo GPU (L40S)0,00117 USD per secondo GPU
  • Fatturato in base al tempo GPU effettivamente utilizzato. All'avvio, 3× il prezzo tipico viene riservato dal tuo saldo e regolato successivamente.
  • 1 credito = 0,01 USD

Calcolatore di costi

Calcolatore prezzi

s

Tipico secondo il provider: circa 11,3 s

Totale

1,32 USD

132 crediti

Per esecuzione

0,0132 USD · 1,32 crediti

Fatturato in base al tempo GPU effettivo; questo è una stima.

05

API

Chiama Dots OCR con la tua chiave API Railwail. Usa questo ID modello nella richiesta:

Nessun esempio API verificato

L'API pubblica passa un formato di input diverso da quello richiesto da questo modello. Usa il playground sopra.

06

Specifiche

ID modello
dots-ocr
Sviluppatore
Community
Categoria
Multimodale
Input
Testo, Immagine
Output
Testo
Fatturazione
In base all'utilizzo (token o tempo GPU)
Voce di catalogo aggiornata
23 settembre 2026

Etichette

  • replicate
  • ocr
  • vision-understanding
  • open-source
07

Domande frequenti

Cos'è Dots OCR?

Dots OCR è un modello di Community nella categoria Multimodale.

Quanto costa Dots OCR su Railwail?

Su Railwail, Dots OCR costa ≈ 0,0132 USD per esecuzione. Ti viene addebitato ciò che ogni richiesta utilizza effettivamente. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.

Quanto è veloce Dots OCR?

Non ci sono ancora abbastanza esecuzioni misurate di Dots OCR su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.

Dots OCR è migliore di BLIP?

Dipende dall'attività. Dots OCR (Community) e BLIP (Salesforce) sono entrambi modelli nella categoria Multimodale. La pagina di confronto mostra i loro prezzi e le specifiche affiancati.

Confronta Dots OCR e BLIP

Dots OCR può elaborare immagini?

Sì. Dots OCR accetta immagini come input oltre al testo.

08

Modelli comparabili

Tutti in questa categoria
  • BLIPSalesforce

    Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.

    ≈ 0,00030 USD/esecuzione

    98 % più economico per unità

    Confronta Dots OCR e BLIP
  • pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.

    ≈ 0,0457 USD/esecuzione

    246 % più costoso per unità

    Confronta Dots OCR e CLIP Interrogator
  • Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.

    ≈ 0,0050 USD/esecuzione

    62 % più economico per unità

    Confronta Dots OCR e Depth Anything v2

Tutti i modelli tramite un'API

Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01 USD.