<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Parser Blog</title>
    <link>https://parserai.co/blog</link>
    <atom:link href="https://parserai.co/blog/feed.xml" rel="self" type="application/rss+xml" />
    <description>Intelligent document processing, OCR and extraction automation.</description>
    <language>en-us</language>
    <item>
      <title>Vendor Statement Reconciliation: How to Match a Supplier Statement Against Your AP Ledger Without Ticking Lines by Hand</title>
      <link>https://parserai.co/blog/vendor-statement-reconciliation-supplier-ledger-accounts-payable</link>
      <guid isPermaLink="true">https://parserai.co/blog/vendor-statement-reconciliation-supplier-ledger-accounts-payable</guid>
      <description>Supplier statements never match your AP ledger on the first pass. See why balances drift, how to classify each difference, and how to automate the match.</description>
      <pubDate>Sat, 26 Sep 2026 13:00:00 GMT</pubDate>
    </item>
    <item>
      <title>One PDF, Twelve Documents: Splitting and Classifying a Batch Before You Extract</title>
      <link>https://parserai.co/blog/document-splitting-classification-mixed-pdf-batches</link>
      <guid isPermaLink="true">https://parserai.co/blog/document-splitting-classification-mixed-pdf-batches</guid>
      <description>Real uploads are not single documents — they are 40-page bundles holding invoices, receipts, contracts and a fax cover sheet. Here is how to find the boundaries, classify each piece, and keep the split auditable before extraction ever runs.</description>
      <pubDate>Thu, 17 Sep 2026 13:00:00 GMT</pubDate>
    </item>
    <item>
      <title>Schema Design for Document Extraction: The Hour That Decides Your Accuracy</title>
      <link>https://parserai.co/blog/extraction-schema-design-field-types-versioning</link>
      <guid isPermaLink="true">https://parserai.co/blog/extraction-schema-design-field-types-versioning</guid>
      <description>Most extraction projects are debugged at the model when the defect is in the schema. Here is how to design the JSON contract — field names, types, enums, nulls, nesting, provenance and versioning — so the model has one obvious right answer for every field.</description>
      <pubDate>Sun, 13 Sep 2026 13:00:00 GMT</pubDate>
    </item>
    <item>
      <title>Prompt Injection Through Documents: How to Harden an LLM Extraction Pipeline</title>
      <link>https://parserai.co/blog/prompt-injection-document-extraction-pipeline-defense</link>
      <guid isPermaLink="true">https://parserai.co/blog/prompt-injection-document-extraction-pipeline-defense</guid>
      <description>An invoice can carry instructions as easily as it carries a total. Here is how indirect prompt injection reaches an extraction pipeline, why instructing the model to ignore it does not work, and the architecture that keeps a hostile PDF from changing what your systems do.</description>
      <pubDate>Sat, 12 Sep 2026 13:00:00 GMT</pubDate>
    </item>
    <item>
      <title>Document Extraction Accuracy: Build a Field-Level Eval Set Before You Trust Any Model</title>
      <link>https://parserai.co/blog/document-extraction-accuracy-field-level-eval-set</link>
      <guid isPermaLink="true">https://parserai.co/blog/document-extraction-accuracy-field-level-eval-set</guid>
      <description>A single “99% accurate” number tells you nothing about whether an extraction pipeline is safe to automate. Here is how to build a field-level eval set, pick the right metrics, and turn accuracy into a routing decision instead of a marketing claim.</description>
      <pubDate>Fri, 11 Sep 2026 13:00:00 GMT</pubDate>
    </item>
    <item>
      <title>How to Integrate Automatic Document Extraction into Your System via API (A Practical Guide for Technical Teams)</title>
      <link>https://parserai.co/blog/how-to-integrate-automatic-document-extraction-into-your-system-via-api-a-practical-guide-for-technical-teams</link>
      <guid isPermaLink="true">https://parserai.co/blog/how-to-integrate-automatic-document-extraction-into-your-system-via-api-a-practical-guide-for-technical-teams</guid>
      <description>A practical, developer-focused guide to integrating automatic document extraction via API: upload, process, poll or webhook, and consume structured JSON — with auth, retries, and idempotency done right.</description>
      <pubDate>Wed, 25 Mar 2026 18:36:23 GMT</pubDate>
    </item>
    <item>
      <title>JWT, IAM, and Auth0: Authentication and Authorization Explained (Without the Confusion)</title>
      <link>https://parserai.co/blog/jwt-iam-and-auth0-authentication-and-authorization-explained-without-the-confusion</link>
      <guid isPermaLink="true">https://parserai.co/blog/jwt-iam-and-auth0-authentication-and-authorization-explained-without-the-confusion</guid>
      <description>Demystify JWT, IAM, and Auth0 — learn authentication vs. authorization, roles, scopes, and secure access for apps, APIs, and microservices.</description>
      <pubDate>Tue, 24 Feb 2026 15:17:29 GMT</pubDate>
    </item>
    <item>
      <title>Markdown vs. JSON vs. Raw Text: How to Optimize Token Usage and Model Performance with Structured Parsing</title>
      <link>https://parserai.co/blog/markdown-vs-json-vs-raw-text-how-to-optimize-token-usage-and-model-performance-with-structured-parsing</link>
      <guid isPermaLink="true">https://parserai.co/blog/markdown-vs-json-vs-raw-text-how-to-optimize-token-usage-and-model-performance-with-structured-parsing</guid>
      <description>Markdown, JSON or raw text? See how each format changes token cost, parse reliability and extraction accuracy — and which to send an LLM for each job.</description>
      <pubDate>Wed, 18 Feb 2026 16:48:20 GMT</pubDate>
    </item>
    <item>
      <title>Beyond OCR: Why Layout-Aware Parsing Is the Secret to High-Precision RAG</title>
      <link>https://parserai.co/blog/beyond-ocr-why-layout-aware-parsing-is-the-secret-to-highprecision-rag</link>
      <guid isPermaLink="true">https://parserai.co/blog/beyond-ocr-why-layout-aware-parsing-is-the-secret-to-highprecision-rag</guid>
      <description>Boost RAG accuracy with layout-aware parsing — not just OCR. Extract structured data from scanned PDFs, invoices, and contracts for high-precision retrieval.</description>
      <pubDate>Fri, 13 Feb 2026 17:39:31 GMT</pubDate>
    </item>
    <item>
      <title>OCR Extraction in 2026: How to Automate Document Processing for Faster, More Accurate Workflows</title>
      <link>https://parserai.co/blog/ocr-extraction-in-2026-how-to-automate-document-processing-for-faster-more-accurate-workflows</link>
      <guid isPermaLink="true">https://parserai.co/blog/ocr-extraction-in-2026-how-to-automate-document-processing-for-faster-more-accurate-workflows</guid>
      <description>Automate OCR extraction in 2026 with intelligent document processing (IDP) to capture invoices, receipts, and contracts faster with fewer errors.</description>
      <pubDate>Tue, 03 Feb 2026 16:34:44 GMT</pubDate>
    </item>
    <item>
      <title>Real Estate Automation in 2026</title>
      <link>https://parserai.co/blog/real-estate-automation-in-2026</link>
      <guid isPermaLink="true">https://parserai.co/blog/real-estate-automation-in-2026</guid>
      <description>Leases, contracts, IDs and income proofs still get read and retyped by hand — here is how real estate teams turn that paperwork into structured data.</description>
      <pubDate>Fri, 30 Jan 2026 20:26:18 GMT</pubDate>
    </item>
    <item>
      <title>Meet Parser: Turn Messy Documents into Clean, Structured Data — Automatically</title>
      <link>https://parserai.co/blog/meet-overview-parser-turn-messy-documents-into-clean-structured-data-automatically</link>
      <guid isPermaLink="true">https://parserai.co/blog/meet-overview-parser-turn-messy-documents-into-clean-structured-data-automatically</guid>
      <description>Parser is an intelligent document processing tool that combines AI understanding with OCR to turn invoices, receipts, contracts, and IDs into clean, structured data — no manual data entry.</description>
      <pubDate>Fri, 16 Jan 2026 15:51:56 GMT</pubDate>
    </item>
  </channel>
</rss>
