# Upstage — Extract structured data from any document—Information Extract API is live

- Company: Upstage (upstage.ai)
- Announced: 2025-05-29
- Category: not stated
- Coverage: not counted
- Announcement: yes
- Group: announcements
- Source: https://www.upstage.ai/blog/en/extract-structured-data-from-any-document--information-extract-api-is-live
- Record: https://forck.live/items/18468-extract-structured-data-from-any-document-information-extract-api-is-live
- Subject: Solar Pro / Solar Mini

One month ago, we opened a playground for Information Extract . The waitlist filled fast, and developers tested it with real-world documents—insurance packets, scanned forms, multipage tables. This early traffic helped us refine schema alignment, layout handling, and batch performance. Today, Information Extract becomes a production-ready REST API —turning unstructured PDFs into structured, schema-aware JSON. No training. No templates. No prompt tuning. What makes Information Extract different Zero-training extraction : Works on any document—no templates, no fine-tuning required Schema-aligned output : Returns structured JSON that matches your schema—types, nesting, and required fields included Layout understanding : Accurately handles tables, checkboxes, multi-page layouts, and rotated content Flat per-page pricing : Predictable billing, regardless of token count or content complexity From document to JSON—in one call Information Extract turns layout-heavy PDFs into clean, typed JSON—aligned to your schema, without templates or scripting. In this example, a multi-page rent roll PDF is converted into structured JSON. …

---

Record: https://forck.live/items/18468-extract-structured-data-from-any-document-information-extract-api-is-live
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
