'Invalid font name' error when extracting text from certain PDF

Hi,

When using the Parser to read all text of certain PDFs, we’re getting the following error:

System.ArgumentException: ‘Invalid font name’

We were using Perser from GroupDocs.Total.NetFramework 26.6 with that code:

            using (var parser = new Parser(documentPath))
            {
                using (TextReader reader = parser.GetText())
                {
                    string strText = reader.ReadToEnd();
            ...

Can you please give it a try with the attached files:
626 WR-Docs - 51097-1.pdf (218.7 KB)
596 WR-Docs - 51231-1.pdf (162.6 KB)
618 WR-Docs - 51231-1.pdf (162.7 KB)

Thanks!

Best regards,
Clemens

Hi @Clemens!

Thank you for reporting this issue. We reproduced this issue and made fixes to deliver in the next October Release

The issue comes from the invalid PDF page content generated by the legacy Gnostice PDFtoolkit.
Strictly, such pages are not valid PDFs, because a font name must resolve in the page’s own resources.
Thank you!

1 Like

@Clemens
We have opened the following new ticket(s) in our internal issue tracking system and will deliver their fixes according to the terms mentioned in Free Support Policies.

Issue ID(s): PARSERNET-2928

You can obtain Paid Support Services if you need support on a priority basis, along with the direct access to our Paid Support management team.

1 Like