Rank: Member
Groups: Registered
Joined: 6/1/2016(UTC) Posts: 28  Location: Hessen Thanks: 2 times Was thanked: 1 time(s) in 1 post(s)
|
I've got incorrect text-extraction information in my document (  doc.zip (398kb) downloaded 23 time(s).). For example the BoundingBox of the first PdfTextObject has a height of 1, which is incorrect according to the fontsize. The first char-rect (GetCharRect(index)) of the mentioned PdfTextObject returns a width and height of 0. But the char-width (GetCharWidth(index)) is 9.329599. These incorrect information are all over the document. Checking the document with Adobe Reader the text-extraction is correct. Can you please have a look on this bug. Thanks and best regards
|
|
|
|
|
|
Rank: Member
Groups: Registered
Joined: 6/1/2016(UTC) Posts: 28  Location: Hessen Thanks: 2 times Was thanked: 1 time(s) in 1 post(s)
|
Is there any solution for that issue?
|
|
|
|
|
|
Rank: Administration
Groups: Administrators
Joined: 1/5/2016(UTC) Posts: 1,138
Thanks: 10 times Was thanked: 133 time(s) in 130 post(s)
|
I have checked this issue and found the problem inside a Pdfium engine. Unfortunately we can't fix it at this moment. I hope this issue will be fixed in the future releases. Currently I can suggest to use the PdfText class instead of PdfTextObject to extract text and its bounding boxes.
|
|
|
|
|
|
Rank: Member
Groups: Registered
Joined: 6/1/2016(UTC) Posts: 28  Location: Hessen Thanks: 2 times Was thanked: 1 time(s) in 1 post(s)
|
Ok thanks, I will have a look.
|
|
|
|
|
|
Forum Jump
You cannot post new topics in this forum.
You cannot reply to topics in this forum.
You cannot delete your posts in this forum.
You cannot edit your posts in this forum.
You cannot create polls in this forum.
You cannot vote in polls in this forum.
Important Information:
The Patagames Software Support Forum uses cookies. By continuing to browse this site, you are agreeing to our use of cookies.
More Details
Close