microsoft/Florence-2-base
Primitive: /extract · Extract ·
Florence-2
This Hub repository contains a HuggingFace's `transformers` implementation of Florence-2 model from Microsoft.
MultimodalText regions
Overview
Hardware: — drives latency, throughput & cost
| Size | 232M params |
|---|---|
| Tasks | /extract |
| License | mit |
| Latency | — |
| Throughput | — |
| Cost | — /1M tok |
Cost is approximate — computed from list GPU prices; your actual price depends on the provider you deploy SIE with.
Extraction
| Output kinds | text_regions |
|---|---|
| Inputs | image |
| Max sequence length | — |
Benchmarks
CORD v2
Key information extraction from receipt images
Corpus: 100 Queries: 100
Quality
f1 0.0000
Compare (0)Compare →