Logo Questions Linux Laravel Mysql Ubuntu Git Menu
 

Parsing linearized pdf xref table

Tags:

parsing

pdf

I have pdf having start as:

%PDF-1.7
%‚„œ”

69 0 obj
<</Linearized 1/L 3937432/O 71/E 2811072/N 9/T 3935937/H [ 996 498]>>
endobj

xref
69 35
0000000016 00000 n

0000001494 00000 n

0000001593 00000 n

0000002065 00000 n

........................

and at the end I have:

0003929147 00000 n

0003929283 00000 n

0003929352 00000 n

0003929458 00000 n

0003935743 00000 n

trailer
<</Size 69/ID[<00E23EA222C14F40B1305A98D798C27F><F53AB532FC064AB39459DBD6BAF21DD6>]>>
startxref
11

Now if I tried to fetch startxref at 11 then I get to „œ” string...which seems wrong, how do I go to actual xrefstart ("xref"), Can Any body help?

like image 530
Mayur Kothawade Avatar asked Aug 23 '26 05:08

Mayur Kothawade


1 Answers

Your xref table byte offset is wrong. It is not 11.

startxref
11

If you fixed this, you can access the xref reading properly

like image 128
suda java Avatar answered Aug 26 '26 23:08

suda java



Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!