Spire.PDF is a professional PDF library applied to creating, writing, editing, handling and reading PDF files without any external dependencies. Get free and professional technical support for Spire.PDF for .NET, Java, Android, C++, Python.
Tue Jul 13, 2021 7:04 pm
Hi all,
i try to extract text of a pdf-File by using SimpleTextExtractionStrategy.
Sadly my extracted Text is looking like this:
"0012300\r\n0045600\r\n007400\r\n007800\r\n0079\n\v31494\f\r\u000e\u000f\u0010\r844\u001100"
is there any chance to convert this to readable text?
thank you
-

strempel
-
- Posts: 1
- Joined: Thu May 27, 2021 8:07 pm
Wed Jul 14, 2021 11:39 am
Hello,
Thanks for your inquiry.
I simulated a PDF file and tested it using the same method as yours. The extracted text is displayed normally. The version I used is the latest Spire.PDF(
Spire.PDF Pack(Hot Fix) Version:7.6.15), if you were not using the latest version, I recommend you give this version a try. If the problem still exists, please provide your sample PDF file for our further investigation. You could attach it here or send to us via email (
[email protected]). Thanks in advance.
Sincerely,
Annika
E-iceblue support team
-


Annika.Zhou
-
- Posts: 1657
- Joined: Wed Apr 07, 2021 2:50 am
Mon Aug 02, 2021 9:20 am
Hello,
Hope you're doing well.
How is your issue going? Has it been resolved? If not, could you please provide your sample PDF file to us for further investigation? You could attach it here or send to us via email (
[email protected]). Thanks in advance.
Sincerely,
Annika
E-iceblue support team
-


Annika.Zhou
-
- Posts: 1657
- Joined: Wed Apr 07, 2021 2:50 am