Extract the field-label and value pairs, checkboxes, sections, signatures, and attachments from any completed or filled-in form PDF. The catch-all extractor.
The full guide: Label-value pairs and checkboxes from a filled form
Need to scale? Create an API key, then run this from your own code or an agent.
Create an API keyFree account required
Upload a file or pick a sample, and see the fields come back.
No signup · nothing stored
The form number, title, and version, the document date, the submitter name, contact email, phone, and address, the signature and signature date, the form status (enum, such as draft, submitted, or approved), and whether the form is an amendment (with the original document number it amends).
Use it as the catch-all when no specific tool fits the form you have. It works on any completed form — a membership form, an intake sheet, a survey, a government or internal form — and returns the field-label and value pairs it finds. If a dedicated tool exists for your document (such as the I-9, benefits enrollment, or notarized document extractor), that one will be more tailored.
The checkboxes selected (array), the field values (array of label and value pairs), and the section responses (array) come back as fields, and a form fields table returns each field id, name, section, type (enum, such as text, date, or checkbox), value, and whether it is required and filled.
Yes. A form sections table returns each section id, title, order, and whether it is complete; a checkboxes table returns each checkbox id, label, whether it is checked, and its section; and a form attachments table returns each attachment id, name, type (enum), date, and size.
The witness name and whether a witness signature is present, whether the form is notarized, and the authorized representative each come back as fields, alongside the form language, effective and expiration dates, and any processing fee. PDF only, up to 10MB and 100 pages; processed via the Talonic API for extraction only and not retained.
The tool reads a single document. The platform reads the whole estate once and keeps it queryable — the same engine, with a memory.
See PDF to Markdown if you run this for data and platform teams, or the extraction API if you are building it in.