Dots OCR
dots-ocrRednote Hilab Dots OCR. End-to-end document parsing model with layout, text and reading-order prediction in one transformer.
- Prezzo
- ≈ 0,0132 USD/esecuzione
- Input → output
- Testo + Immagine → Testo
- Sviluppatore
- Community
- Aggiornato
- 23 settembre 2026
Playground
Prova Dots OCR
Nessuna maschera di input
Nessun modulo di input per questo modello
I suoi input non sono ancora documentati. Per evitare errori di esecuzione, non offriamo un modulo qui. Scegli invece un modello comparabile.
Examples
InputPrompt
grounding_ocr
Response
[{"bbox": [448, 513, 528, 535], "category": "Section-header", "text": "### Directions"}, {"bbox": [448, 533, 824, 565], "category": "Text", "text": "For sunscreen use: • apply generously 15 minutes before sun exposure • reapply: • after 80 minutes of swimming or sweating • immediately after towel drying • at least every 2 hours"}, {"bbox": [448, 563, 832, 624], "category": "List-item", "text": "• **Sun Protection Measures.** Spending time in the sun increases your risk of skin cancer and early skin aging. To decrease this risk, regularly use a sunscreen with a Broad Spectrum SPF value of 15 or higher and other sun protection measures including: • limit time in the sun, especially from 10 a.m. - 2 p.m. • wear long-sleeved shirts, pants, hats and sunglasses"}, {"bbox": [448, 622, 640, 638], "category": "List-item", "text": "• children under 6 months of age: Ask a doctor"}]
Informazioni su Dots OCR
Dots OCR è un modello di Community nella categoria Multimodale. Su Railwail, Dots OCR costa ≈ 0,0132 USD per esecuzione.
Prezzi
| Esecuzione tipica (≈ 11 s su L40S) | 0,0132 USD per esecuzione |
|---|---|
| Tempo GPU (L40S) | 0,00117Â USD per secondo GPU |
- Fatturato in base al tempo GPU effettivamente utilizzato. All'avvio, 3× il prezzo tipico viene riservato dal tuo saldo e regolato successivamente.
- 1 credito = 0,01Â USD
Calcolatore di costi
Calcolatore prezzi
Tipico secondo il provider: circa 11,3 s
Totale
1,32Â USD
132 crediti
Per esecuzione
0,0132 USD · 1,32 crediti
Fatturato in base al tempo GPU effettivo; questo è una stima.
API
Nessun esempio API verificato
L'API pubblica passa un formato di input diverso da quello richiesto da questo modello. Usa il playground sopra.
Specifiche
- ID modello
dots-ocr- Sviluppatore
- Community
- Categoria
- Multimodale
- Input
- Testo, Immagine
- Output
- Testo
- Fatturazione
- In base all'utilizzo (token o tempo GPU)
- Voce di catalogo aggiornata
- 23 settembre 2026
Etichette
- replicate
- ocr
- vision-understanding
- open-source
Domande frequenti
Cos'è Dots OCR?
Dots OCR è un modello di Community nella categoria Multimodale.
Quanto costa Dots OCR su Railwail?
Su Railwail, Dots OCR costa ≈ 0,0132 USD per esecuzione. Ti viene addebitato ciò che ogni richiesta utilizza effettivamente. L'utilizzo viene pagato con crediti prepagati; 1 credito equivale a 0,01 USD.
Quanto è veloce Dots OCR?
Non ci sono ancora abbastanza esecuzioni misurate di Dots OCR su Railwail per indicare un tempo di esecuzione. Dipende dall'input, dalle impostazioni e dal carico presso il provider.
Dots OCR è migliore di BLIP?
Dipende dall'attività . Dots OCR (Community) e BLIP (Salesforce) sono entrambi modelli nella categoria Multimodale. La pagina di confronto mostra i loro prezzi e le specifiche affiancati.
Confronta Dots OCR e BLIPDots OCR può elaborare immagini?
Sì. Dots OCR accetta immagini come input oltre al testo.
Modelli comparabili
Tutti in questa categoria- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- CLIP InterrogatorCommunity
pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
- Depth Anything v2Community
Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
Tutti i modelli tramite un'API
Una chiave API per tutti i modelli su Railwail. L'utilizzo viene addebitato da crediti prepagati, 1 credito = 0,01Â USD.