.xls saved with an .xlsx name still opens. PDF and Word use an extract that is already ready. This step does not start one. Call Extract tables first.
How it runs
1
Resolve File id
One opaque id. Same rules as Extract tables.
2
Outline or one sheet
Empty Sheet lists every sheet: name, clipped rows and columns, how many rows have text, how many merged ranges. A sheet name returns that sheet’s map, or its cells when Rows is set.
3
Compress the sheet
Rows with the same filled columns become one line (
from-to). The line shows those columns (cols=A-D,F), the first values of its first row with their column letters, and the merged ranges that start in it. A header, a group row, or a row with a different set of filled columns stays its own line.4
Fold repeated blocks
When the same pattern of lines repeats three times or more, for example one table per carrier or per port, the copies become one line.
6-17 repeats 1-5 x3 means rows 6 to 17 hold three more copies of the pattern in rows 1 to 5. Then it lists where each copy starts and its first values. A copy can have more or fewer rows than the first one.Input
Inspector Settings. Empty-field rules for the Graph: Previous nodes.
Merged text stays on the top-left cell of the range. Other cells in the range are empty. The map names the range; it does not copy the text across it.
Empty trailing columns that the file marks as used are dropped. A sheet with no text fails. A sheet over 8000 rows, 64 columns, or 8000 merged ranges fails when you name it, and the error includes the counts. The outline still lists that sheet, and Extract tables still reads it.
PDF and Word add
warning: merges_unavailable. Merged cells from those files are not listed.
Output
On the next node, Previous nodes listsSheet map: Result.
A long map is stored on the run. The model reads the rest with the result id from the tool result.
A map:
2-4: