THausherr commented on PR #526:
URL: https://github.com/apache/pdfbox/pull/526#issuecomment-5848377791

   Here's a correction for the awt test. It's not great but it does the job:
   
   ```
               PDFTextStripper stripper = new PDFTextStripper();
               String s = stripper.getText(doc);
               String sStripped = s.replaceAll("\r", "").replaceAll(" +"," ")
                       .replaceAll(" *\\n", "\n")
                       .strip();
   
               String text =
                       FIRACODE_STRING + "\n" + 
                       FIRACODE_STRING + " (Ligatures)" + "\n" + 
                       DEJAVU_STRING + "\n" + 
                       DEJAVU_STRING + " (Ligatures)" + "\n" + 
                       DEJAVU_STRING + " (Kerning)" + "\n" + 
                       DEJAVU_STRING + " (Ligatures and kerning)" + "\n" + 
                       THAI_STRING + "\n" + 
                       BENGALI_STRING + " (ভারত)"+ "\n" + 
                       BENGALI_STRING2 + " " + BENGALI_STRING2;
               if (useActualText)
               {
                   assertEquals(text, sStripped, "Extracted Text should equal 
the written text for " + outputPDFFilename);
               }
   ```


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]


---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to