Connect with us

Technology

Nutrient Data Extraction API launches for source-grounded, production-ready document AI

Published

on

Nutrient Data Extraction transforms PDFs, scans, images, and Office files into structured document data — parsing content into spatial JSON or Markdown, and extracting schema-defined fields with per-field confidence signals and citations back to the source, helping enterprises operationalize agentic document workflows with reliable structured data.

RALEIGH, N.C., Sept. 9, 2026 /PRNewswire/ — Nutrient Data Extraction API, a document parsing and structured data extraction service, is now generally available for enterprises building AI agents, RAG systems, large corpus document query, and enterprise automation. Nutrient Data Extraction helps organizations operationalize AI by transforming complex documents into reliable structured data for production-ready agentic workflows. 

Nutrient Data Extraction parses PDFs, scans, images, and Office files into spatial JSON or Markdown, and extracts schema-defined fields with per-field confidence signals and source citations, giving enterprises accurate structured data from documents enabling production-capable agentic workflows. 

Nutrient Data Extraction at a glance:

Inputs: PDFs, scans, images, and Office files

Parse outputs: spatial JSON or Markdown

Modes: text, structure, understand, and agentic

Structured extraction: customer-defined JSON Schema

Extraction grounding: page references, bounding boxes, source blocks, match labels, and confidence signals

Routing: fuzzy_match and not_found can go to human review

Access: REST API, free tier, Studio, and getting-started guide

Languages: 100+ for OCR

Security: HTTPS/TLS encryption in transit, SOC 2 Type 2 audited infrastructure 

The reliability gap in agentic document workflows

AI has made document intelligence possible at enterprise scale. Documents can now be read, summarized, classified, extracted, and acted on in seconds. 

But generic LLM extraction is inherently probabilistic: the same document can produce different outputs under different conditions, lose structure, and return values with no traceable link back to the page they came from. Teams either accept the risk of unverified data flowing into ERPs, claims systems, and approval workflows, or they revert to manual review. 

Organizations heavily need data that is reliable, explainable, and traceable back to its source for agentic workflows to be trustworthy and ready for production. 

That’s the reliability gap Nutrient Data Extraction was built to close. Most LLMs can extract data from documents. Far fewer systems can extract it accurately, consistently, and explain where it came from at a quality level that stands up in an audit.

That’s the difference between model output alone and source-grounded document infrastructure built for production-ready agents. It’s the difference between AI that’s impressive in a demo and AI that can be relied upon by the human signing off on the business outcome.

And because Nutrient Data Extraction is a foundational part of Nutrient’s broader document platform, teams can move from parsing and extraction to unlocking routing, governance, human review, and execution all without downstream workflow automation losing the auditable link back to the source document.

Built for multiple levels of understanding.

Teams, including Nutrient’s implementation teams, are able to customize the output format per request based on the type and complexity of the document and the system requirements whether that’s spatial JSON for layout-aware elements with reading order, bounding boxes, tables, and key-value regions, or Markdown for RAG, search indexing, and general document Q&A. Nutrient Data Extraction detects tables, forms, formulas, charts, handwriting, checkboxes, and headings, and supports more than 100 different OCR languages.

Not every document requires the same level of analysis and reasoning to get to accurate structured data. Nutrient Data Extraction lets teams optimize for speed, cost, or depth through four processing modes: text, structure, understand, and agentic, allowing processing to scale only when additional understanding is required.

Every extracted value traces back to the source

The extract endpoint maps a document to a customer-defined or intelligently generated JSON Schema and returns each requested field with a bounding box, page reference, the source blocks it was drawn from, a confidence signal, and a label describing how the value was grounded. Source grounding can drive workflow decisions automatically. Fields labeled “fuzzy_match” or “not_found” can be routed to human review, while high-confidence matches continue directly into downstream workflows.

Unlike tools that present confidence as a probability, Nutrient Data Extraction treats confidence as a relative signal. That distinction helps organizations structure and make better workflow decisions instead of over-interpreting a single score.

“Most agents can pull data out of a document. Almost none of them can prove where the data came from,” said Jonathan Rhyne, co-founder and CEO of Nutrient. “Every field we return points back to the exact spot in the source, and when we can’t find something, we say so instead of guessing. In production, ‘Trust us’ isn’t an answer enterprises can take to their auditors or something accountable parties will rely on.”

Published accuracy benchmarks

Nutrient publishes open accuracy benchmarks and updates the published results as new versions ship. In our latest test run, understand mode scored 0.932 overall accuracy on the publicly available 200-document opendataloader-bench corpus, across reading order, table structure, and heading hierarchy. Text and structure modes are scored on the public leaderboard, while understand and agentic modes were evaluated internally against the same corpus. The original open-source comparison was run on July 6, 2026 and the current results and methodology are maintained at nutrient.io/api/data-extraction-api/benchmarks

Separately, Nutrient has published two open grounding artifacts on Hugging Face: grounding-en, a model that scores whether an extracted value is supported by evidence in the source document, released under Apache-2.0; and the grounding benchmark dataset it is evaluated against, released under CC-BY-SA-4.0.

What it handles in practice

Nutrient publishes interactive extraction demos with no signup or API key required, where hovering an extracted field highlights the exact region of the source document it came from:

Healthcare: Table-row extraction from a CMS-1500 claim form, including the dropout-red ink grid that defeats most scanners.

Government and vital records: Signature block detection on a mixed handwritten and printed birth record, separating a printed name from the adjacent cursive signature.

Mortgage and financial services: Value isolation in a dense three-column financial table where Monthly Income, Total Assets, and Total Expenses sit side by side under near-identical labels.

Legal and contracts: Full multi-sentence narrative extraction from a free-text field in a small claims filing, plus cross-page extraction across three pages.

AI, search, and RAG: An appraisal report decomposed into semantic blocks with spatial coordinates, so a retrieval pipeline can cite its exact source.

Availability

Nutrient Data Extraction is generally available via a REST API or can be built into custom solutions by our implementation teams. New accounts receive 5,000 Data Extraction API credits per month at no cost, with no credit card required. Signup is agent-compatible. 

Developers and agents can send a first request in minutes using the getting-started guide, or open the interactive demos in a browser with no account. Nutrient Data Extraction Studio provides a visual testing environment for evaluating documents before writing any code. 

All API communication uses HTTPS/TLS encryption, and the platform is SOC 2 Type 2 audited, with reports available under NDA. Processing-run retention varies by plan. Customers can delete runs and runs expire after the plan-specific retention period. On plans without internal data retention, documents are explicitly not retained for service improvement or model training. 

To view documentation, review the published benchmarks, or get a free API key, visit nutrient.io/api/data-extraction-api.

About Nutrient

Nutrient is the deterministic document platform organizations rely on to operationalize production capable agentic systems for document centric workflows. It combines reliable document processing infrastructure for agents, intelligent routing and governance, and interfaces for the human in the loop. Built on a decade of enterprise document expertise, Nutrient provides agentic document workflow solutions for thousands of organizations worldwide, including more than 15 percent of the Global 500, thousands of commercial businesses across 80 countries, and more than 130 public sector organizations in 24 countries. Backed by Insight Partners, Nutrient is headquartered in Raleigh, North Carolina, with offices in England, France, and Austria. Learn more at nutrient.io.

View original content to download multimedia:https://www.prnewswire.com/news-releases/nutrient-data-extraction-api-launches-for-source-grounded-production-ready-document-ai-302873992.html

SOURCE Nutrient

Continue Reading

Technology

proof-it and The Tennessee Whiskey Trail Announce The 2026 TN Collective Tasting Notes Contest

Published

on

By

ATLANTA, Sept. 11, 2026 /PRNewswire/ — The 2026 Tennessee Collective, a collaborative effort between 11 Tennessee distilleries, and proof-it, a next-generation consumer authenticity and engagement platform, proudly announces a partnership to commemorate the release of the The 2026 Tennessee Collective American Whiskey, a delicately crafted new addition to the whiskey industry of Tennessee that will give consumers a unique look into the craft, process, and flavor of each bottle, and direct participation with those who made it.

Each bottle will come with proof-it technology, giving consumers the chance to see the passion in each bottle.

A celebration of the long and rich history of distilling in the state of Tennessee, this new American Whiskey from The Tennessee Collective is a bold testament to the dedication and alliance of the industry. WIth efforts from Big Machine Distillery, Gate 11 Distillery, Jack Daniel’s Distillery, Leiper’s Fork Distillery, Nelson’s Green Brier Distillery, Old Dominick Distillery, Old Glory Distilling Co., Peg Leg Porker Distillery, Short Mountain Distillery, Sugarlands Distilling Co., and Tennessee Legend Distillery, this bottle is a new entry in the history and camaraderie of the Tennessee whiskey distilling industry.

“proof-it has become an integral partner in the Tennessee Collective, creating an exclusive digital experience that connects consumers more deeply to the bottle, the blend and the distilleries behind it,” said Charity Toombs, Executive Director of the Tennessee Whiskey Trail. “Just as importantly, their technology has given us a valuable window into our customers—helping us better understand who they are, how they engage with the Collective and what interests them most about the product and blend.”

The whiskey will be available starting September 1, 2026 at each of the participating distillery’s tasting rooms, as well as online for pre-order at www.tnwhiskeytrail.com/tn-collective. Each bottle will come with proof-it tag technology, giving consumers the chance to see the passion that went into each bottle, while also offering the chance to discover what makes this bottle unique. This year’s release unlocks a new way to engage with the story of each distiller involved, as each individual distillery’s allocation of bottles will come with a custom video and the story of their specific contribution to the whiskey, only accessible via the proof-it tag on the neck. Beyond a collector’s item, each bottle tells the story of its own existence in eleven different parts and celebrates the craftsmanship that went into this release.

To celebrate the launch, proof-it will also be hosting the Tennessee Collective Tasting Notes Contest through January 7, 2027, offering buyers an opportunity to explore the whiskey’s depth. Simply tap an unlocked smartphone to the unique proof-it tag on each bottle, and users can offer opinions and notes on what they taste in the rich complexity of this unique whiskey. The best entry to the Tennessee Collective Tasting Notes Contest will win tickets to Grain and Grits Festival 2026!

About Tennessee Whiskey Trail

Where the whiskey meets the road. Seven years ago, the Tennessee Whiskey Trail was launched by the Tennessee Distillers Guild as a celebration of our state and its signature spirits. Today we host more than 30 stops across the state, offering a taste of Tennessee and a road map for adventure. Visitors to the Trail can soak up the sights and sounds that have been shaped by Tennessee whiskey, whether stopping by for a sip or making a weekend of it. For more information, visit www.tnwhiskeytrail.com and the Tennessee Whiskey Trail’s Instagram @tnwhiskeytrail.

About proof-it

The proof-it™ platform enables consumers to learn the stories behind the labels of the brands they love, and manage their collections in revolutionary ways. Combining authenticity, consumer engagement, and tamper-proof technology at once, proof-it™ is leading the way into the next generation of consumer engagement and brand building. Look for the proof-it™ Wings™ logo and simply tap it with your unlocked phone.To learn more about the ways proof-it™ can enhance your consumer experience and beyond, visit www.proofit.app.

proof-it™ is powered by Aware Technologies, Inc., providing complete visibility throughout a product’s journey, and enabling new ways for brands to interact with their customers.

Full contest rules are available here: https://scan.proofit.app/contests/rules/tennessee-collective-rules-2026

View original content to download multimedia:https://www.prnewswire.com/news-releases/proof-it-and-the-tennessee-whiskey-trail-announce-the-2026-tn-collective-tasting-notes-contest-302876770.html

SOURCE Aware Technologies, Inc.

Continue Reading

Technology

IDScan.net Data Breach: Edelson Lechtzin LLP Launches Investigation Into Exposure of Personal Information

Published

on

By

National class action firm offering free case evaluations to individuals impacted by IDScan.net cybersecurity incident

NEW ORLEANS, Sept. 11, 2026 /PRNewswire/ — Edelson Lechtzin LLP, a national class action law firm, is investigating data privacy claims arising from the IDScan.net data breach. IDScan.net experienced a data breach on or about September 1, 2026.

What Happened

IDScan.net learned around September 1, 2026, that an unauthorized party may have accessed or copied customer information stored in its cloud accounts. The company secured its systems and brought in outside cybersecurity specialists to investigate the incident, which remains ongoing.

Information Exposed

The IDScan.net data breach may have compromised certain personal information, including full names and driver’s license or other government-issued identification numbers.

Who May Be Impacted

Individuals who received a data breach notification from IDScan.net may face an increased risk of identity theft and fraud.

Your Legal Options

Edelson Lechtzin LLP is investigating a potential class action to pursue legal remedies for individuals whose sensitive personal data may have been compromised in the IDScan.net breach. The firm will evaluate your rights and potential claims at no cost.

Recommended Protective Steps

Review account statements and credit reports regularly and remain vigilant for suspicious activity. Confirm whether your information was involved in the IDScan.net incident and preserve any letters or emails you received about the breach. Consider placing fraud alerts and credit monitoring.

Contact Us for a Free Case Evaluation

Speak confidentially with a data privacy attorney today: Marc Edelson, Esq., Edelson Lechtzin LLP, 411 S. State Street, Suite N-300, Newtown, PA 18940; Phone: 844-696-7492; Email: medelson@edelson-law.com; Web: www.edelson-law.com. Or click HERE to request a free consultation.

About IDScan.net

IDScan.net is a technology company that provides identity verification and data capture solutions for businesses and organizations.

About Edelson Lechtzin LLP

Edelson Lechtzin LLP is a national class action law firm with offices in Pennsylvania and California. In addition to data breach litigation, the firm handles class and collective actions involving securities and investment fraud, federal antitrust violations, ERISA employee benefit plans, wage theft, and consumer fraud

Media and Partnership Inquiries: Use the contact information above to connect with our team regarding interviews, co-counsel opportunities, and referral partnerships.

Legal Notice: This press release may be considered Attorney Advertising in some jurisdictions.

View original content to download multimedia:https://www.prnewswire.com/news-releases/idscannet-data-breach-edelson-lechtzin-llp-launches-investigation-into-exposure-of-personal-information-302876772.html

SOURCE Edelson Lechtzin LLP

Continue Reading

Technology

Real Time • Tide Watching | “Greater BRICS cooperation has become the leading echelon of the Global South” —- Qian Feng and Xu Feibiao on 20 Years of BRICS Cooperation and the Next Decade

Published

on

By

NEW DELHI, Sept. 11, 2026 /PRNewswire/ — A report from haiwainet.cn

Marking the 20th anniversary of the BRICS cooperation mechanism and ahead of the New Delhi Summit, this episode of Real Time • Tide Watching is joined by Professor Qian Feng, Fellow and Director of the Research Department at Tsinghua University’s National Strategy Institute, and Professor Xu Feibiao, Research Professor and Director of the Center for BRICS and G20 Studies of China Institutes of Contemporary International Relations (CICIR). They focus on the highlights of the 18th BRICS Summit, new progress in BRICS cooperation, and the trajectory of China-India relations, and engage in an in-depth discussion on the next decade of BRICS cooperation.

 

View original content to download multimedia:https://www.prnewswire.com/news-releases/real-time–tide-watching—greater-brics-cooperation-has-become-the-leading-echelon-of-the-global-south–qian-feng-and-xu-feibiao-on-20-years-of-brics-cooperation-and-the-next-decade-302876780.html

SOURCE haiwainet.cn

Continue Reading

Trending