Guide · Inbound

How to turn a PDF purchase list into clean rows without retyping it.

The purchase list arrives as a PDF with merged columns. Someone retypes it into a spreadsheet every cycle, and every typo later shows up as a stock-out.

Most distribution warehouses receive their procurement or purchase list as a PDF from a head office, a cooperative or a supplier. The PDF looks tidy on screen, but copying it out is slow: columns are merged, quantities for several hubs sit in one cell, and product names do not match the names in your own master.

Why copy-paste and generic converters fail

A reliable way to do it

  1. Read every row of the PDF, including merged columns, into a structured table.
  2. Split quantities by hub so each row is one product for one location.
  3. Match each line to its SKU in the product master: packing, weight and storage zone.
  4. Flag mismatches for a person to resolve; never guess a match.
  5. Only then let the rows count as expected stock for receiving at the gate.

A quick quality check before you receive

CheckWhy it matters
Row count equals the PDF’s line countCatches rows lost at page breaks
Total quantity per hub equals the PDF totalsCatches split errors
Every line matched to a SKUUnmatched lines become stock-outs later
Weight known for every SKUNeeded later for bag allocation

How Warehouse OS does it

Warehouse OS is built for distribution warehouses that fulfil many small orders every cycle. The cycle’s purchase list is read from the PDF, including merged columns, and each line is checked against the product master. Receipts are logged with vehicle, gate and challan; stock is put away to the zone and rack a 3D twin of the floor suggests. Member orders are joined with product weights so every order gets a small bag, a large bag or a split before packing starts. Each permanent bag tag is scanned to a durable dispatch box carrying its route, van and member, and every receipt, move and dispatch stays on one record. It runs on web and handheld.

Questions

How do I convert a PDF purchase list into Excel rows?

Read every row including merged columns into a table, split quantities by hub, then match each line against your product master and flag anything that does not match before it counts.

Why do PDF-to-Excel converters give wrong rows?

Purchase lists often merge several fields into one cell and carry quantities for many hubs per line; a generic converter copies the layout, not the meaning, and does not check against your product master.

What should be checked after extracting a purchase list?

Row counts, totals per hub, a SKU match for every line and a known weight for every SKU.

See it on your own data. Warehouse OS — Purchase list to packed orders. Book a 30-minute working session with an engineer.

General guidance, current as of the date above. Figures and examples are illustrative unless a source is linked.