Tutorial: Extracting Unstructured Text Using Large Language Models
This tutorial shows how to build a two-stage pipeline in which a multimodal large language model reads each page as an image and a second model turns the extracted text into structured JSON. We cover the choices that make such pipelines reliable and affordable, including structured outputs, parameter control, task decomposition, cost-effective model selection, and validation against a human-verified ground truth.
Authors
- Oliver Schaer (ORCID: https://orcid.org/0000-0003-1878-8134)
- Simon Spavound
- Panos Markou (ORCID: https://orcid.org/0000-0002-4210-281X)
Institutions
- University of Virginia (US)
- Drexel University (US)
Publication Details
- Journal
- INFORMS Journal on Applied Analytics
- Published
- 2026-10-09
- DOI
- https://doi.org/10.1287/inte.2026.0314
- Primary Topic
- Topic Modeling
- Type
- article
- Field-Weighted Citation Impact
- 0.00