Hewto.ai
Products · CMS 1500 OCR Software for EDI

Convert CMS-1500 to EDI 837P, automatically

Hewto.ai is CMS 1500 OCR software and a healthcare claims OCR API purpose-built for paper claim conversion — turning scanned, handwritten, and red drop-out ink HCFA / CMS-1500 forms into a validated ANSI X12 837 Professional (837P) file. End-to-end paper claim conversion, from image to 837 text file.

See it in action

Watch Hewto.ai process a CMS-1500 in real time

A two-minute walkthrough of ingesting, extracting, and validating a HCFA / CMS-1500 claim — from scan to 837P-ready structured data.

CMS 1500 OCR software

Layout-aware OCR reads faxed, scanned, and photographed CMS-1500 claim forms at any quality — including handwriting and stamps.

Health insurance claim data extractor

Extracts every box of the health insurance claim form (Boxes 1–33): patient, insured, diagnosis pointers, modifiers, and rendering/billing NPIs.

Validation & edits

Applies NPI, ICD-10, and CPT checks plus payer edits with per-field confidence scoring before data leaves the pipeline.

Clean output & APIs

Delivers structured JSON or 837P-ready data to your clearinghouse, EHR, or RCM platform.

Agentic Document Extraction

Why Hewto.ai outperforms traditional CMS-1500 OCR

We don't use character-by-character OCR. Hewto.ai runs an agentic document extraction pipeline built on vision-language models fine-tuned specifically for the HCFA / CMS-1500 form. The model reads the claim the way a trained biller does — looking at the whole form, its layout, and every visual cue — so it captures the details a traditional OCR engine silently misses.

Vision-first, not character-first

A VLM sees the form as a whole — box borders, checkboxes, tick-marks, signatures, stamps, red drop-out ink, and overlapping handwriting — instead of stitching guesses letter by letter. Visual context stays intact.

Fine-tuned on CMS-1500 forms

Our models are fine-tuned on real HCFA / CMS-1500 claims across payers, scan qualities, and handwriting styles. They know Box 24 is a service line — not just a grid of characters — so extractions map cleanly to fields, not raw text.

Agentic reasoning, per field

Each field runs through a purpose-built agent that verifies context, cross-checks related boxes, and self-corrects when something looks off — the kind of judgment a deterministic OCR pipeline simply can't make.

Traditional OCR misses
  • Checkbox and tick-mark selections in Boxes 1, 10, 11d, 27
  • Handwritten patient signatures and dates
  • Values written outside the printed field lines
  • Red drop-out ink bleeding into extracted text
  • Modifier stacks and pointers linked to the wrong service line
  • Provider stamps, initials, and manual corrections
Hewto.ai captures
  • Every checkbox and tick — including partially marked ones
  • Handwriting in any box, including signatures and dates
  • Out-of-bounds writing correctly assigned to the right field
  • Clean data from red drop-out ink CMS-1500 templates
  • Correct diagnosis pointers → service line relationships
  • Stamps, initials, and hand-corrections preserved with confidence scores
One-click VLM fine-tuning

Fine-tune a vision-language model at the click of a button

Hewto.ai provides vision-language model fine-tuning at the click of a button. Upload a handful of labeled samples of CMS-1500 forms — or any new form variant, payer template, or document type you handle — and Hewto.ai fine-tunes a purpose-built VLM for it. No ML team, no training pipelines, no infrastructure to manage.

  • Upload samples, click fine-tune, ship
  • No ML engineers or training pipelines required
  • Custom accuracy on your document mix
  • Ready to run in minutes, not weeks
Fine-tune workflow
  1. 01Upload labeled samplesDrop in 20–200 examples of your target form.
  2. 02Click Fine-tuneHewto.ai kicks off a fine-tuning run — no config, no code.
  3. 03Ship the modelYour custom VLM goes live behind the same API endpoint.
Full box coverage

Every field on the professional claim

As a complete health insurance claim data extractor, Hewto.ai's CMS 1500 OCR software captures every box of the claim form — so nothing is re-keyed by hand.

  • Boxes 1–13: patient & insured demographics
  • Box 21: ICD-10 diagnosis codes
  • Box 24: service lines, CPT/HCPCS, modifiers, charges
  • Boxes 24J / 31–33: rendering & billing provider NPIs
HCFA / CMS-1500 health insurance claim form
CMS-1500 · extracted
ICD-10E11.9
CPT (24)99213
Charge$185.00
99.4% confidence
Paper claim to 837P converter

Convert HCFA to 837 Professional

Hewto.ai parses CMS-1500 to ANSI 837 and generates a compliant X12 837P file — a true HCFA 1500 to EDI converter. Every box is mapped through a CMS-1500 to EDI 837P crosswalk, so a scanned or faxed claim becomes clean, submittable EDI with no manual keying.

CMS-1500 → X12 loop / segment mapping

  • Box 24 service lines → Loop 2400 (SV1, DTP service dates)
  • Box 21 diagnoses → Loop 2300 HI (health care diagnosis codes)
  • Boxes 1–13 patient & insured → Loops 2010BA / 2010CA (subscriber & patient)
  • Boxes 31–33 rendering & billing provider → Loops 2310B / 2010AA (NM1, PRV, NPI)
  • Charges & totals → CLM segment and monetary (AMT) elements
claim.837 (X12 837P)
ST*837*0001*005010X222A1~
CLM*PATIENT01*185.00***11:B:1~
HI*ABK:E119~
NM1*82*1*SMITH*JANE****XX*1234567890~
LX*1~
SV1*HC:99213*185.00*UN*1***1~
DTP*472*D8*20250114~

Generated automatically from a scanned CMS-1500 image — image to 837 text file.

Digital mailroom & data capture

CMS-1500 data extraction software for real-world claims

CMS-1500 digital mailroom software

Auto-classify and process every CMS-1500 hitting your fax lines and mailroom — automated paper claims digitization for healthcare, end to end.

Red drop-out ink OCR

Reads the red drop-out ink CMS-1500 template and lifts only the entered data — no printed form lines bleeding into results.

Handwritten CMS-1500 extraction

AI handwriting recognition captures hand-filled boxes that defeat traditional OCR.

Scanned HCFA 1500 PDF to 837

Convert scanned HCFA 1500 PDFs and images to 837 format — turning a CMS-1500 image into an 837 text file automatically.

For developers

CMS 1500 parsing API & healthcare claims OCR API

Integrate claim capture directly: an X12 837 generator from paper-claim JSON/XML, and an SDK to convert image to 837P EDI.

CMS 1500 parsing API

A healthcare claims OCR API that parses CMS-1500 to ANSI 837 — POST an image, receive structured JSON/XML in response.

SDK: convert image to 837P EDI

Drop-in SDK to convert a claim image to 837P EDI, backed by an X12 837 generator from paper-claim JSON/XML.

Automated claim intake BPO

Offload high-volume paper claim intake — automated claim intake BPO with human-in-the-loop QA and SLAs.

CMS 1500 OCR software FAQ

Health insurance claim data extraction, answered

Common questions about Hewto.ai's CMS 1500 OCR software and health insurance claim data extractor for the HCFA / CMS-1500 claim form.

Ready to see Hewto.ai on your documents?

Book a personalized demo and watch us process your HCFA, UB-04, and dental claims live.