Spire.PDF is a professional PDF library applied to creating, writing, editing, handling and reading PDF files without any external dependencies. Get free and professional technical support for Spire.PDF for .NET, Java, Android, C++, Python.

Fri Apr 14, 2023 2:28 am

I had try extract table with Ver 9.4.0, but Text not correct:
Ex: Text = "小径油穴付き超硬ドリル 2Dタイプ
Small Diameter Carbide Drill with Internal Coolant Supply
(2D t ype) "

Result =

Code: Select all
小S径a油 き超硬ド ル 2D
m ll Diam穴付eter Carbide Drilリl with Interタイnal Coプolant Supply
(2D type)

Code: Select all
string strTargetPage = @"D:\admin\SpirePDF\p14.pdf";         
            Spire.Pdf.PdfDocument doc = new Spire.Pdf.PdfDocument();           
            doc.LoadFromFile(strTargetPage);
            PdfTableExtractor extractor = new PdfTableExtractor(doc);
            PdfTable[] tableLists = extractor.ExtractTable(0);       
            if (tableLists != null && tableLists.Length > 0)
            {
                int iTable = 0;
                foreach (PdfTable table in tableLists)
                {
                    iTable += 1;
                    int iTotalRow = table.GetRowCount();
                    int iTotalCol = table.GetColumnCount();
                    //Loop though the row and colunm
                    for (int iRow = 0; iRow < iTotalRow; iRow++)
                    {
                        for (int iCol = 0; iCol < iTotalCol; iCol++)
                        {
                            //Get text from the specific cell
                            string text = table.GetText(iRow, iCol);                         

                        }
                    }
                }
            }

Note: i try with Ver 8.1.0, it extract ok.
Sample data: https://drive.google.com/file/d/1ACR5Mb ... share_link

daitranthanhhoa
 
Posts: 51
Joined: Mon Sep 19, 2016 3:04 am

Fri Apr 14, 2023 7:22 am

Hi,

Thank you for your inquiry.

I have used our Spire.PDF for .NET (Version 8.1.0 and Version 9.4.0) to extract table content from your PDF document, and was able to reproduce the issue where the extracted result content is incorrect in Version 9.4.0. I have logged this issue in our issue tracking system with issue tracking number SPIREPDF-5939.

Our development team will further investigate and fix it. Once there are any updates, we will keep you informed. We apologize for any inconvenience caused.

Sincerely,
Ella
E-iceblue support team
User avatar

Ella.Zhang
 
Posts: 42
Joined: Fri Apr 07, 2023 7:42 am

Tue Apr 18, 2023 10:29 am

A Other Table extract missing Text:
refer: https://drive.google.com/file/d/1I82Py8 ... share_link

Text ="Matrix"
But Result ="Matr"

daitranthanhhoa
 
Posts: 51
Joined: Mon Sep 19, 2016 3:04 am

Wed Apr 19, 2023 5:55 am

Hi,

Thank you for your feedback.

I used our Spire.PDF for.NET to extract the table content from your new PDF document and reproduced the problem of incorrect text content in the result.I have updated this issue in our tracking system with the ticket number SPIREPDF-5939.

Our development team will conduct further investigation and fix it as soon as possible. We will keep you updated once we have any updates. We apologize for any inconvenience this may have caused.

Sincerely,
Ella
E-iceblue support team
User avatar

Ella.Zhang
 
Posts: 42
Joined: Fri Apr 07, 2023 7:42 am

Tue May 23, 2023 9:53 am

Hi,

Thanks for your patience.
Glad to inform you that we just released Spire.PDF 9.5.4 hotfix , which has fixed your issue SPIREPDF-5939, please download from the following links and have a test.
Website: https://www.e-iceblue.com/Download/download-pdf-for-net-now.html
Nuget: https://www.nuget.org/packages/Spire.PDF/9.5.4

Best regards,
Triste
E-iceblue support team
User avatar

Triste.Dai
 
Posts: 1000
Joined: Tue Nov 15, 2022 3:59 am

Return to Spire.PDF

cron