I have noticed that character corruption occurs when converting Japanese text files with JIS (ISO-2022-JP) encoding via the "Convert Text Files to PDF" dialog, even when "Japanese (JIS)" is selected in the "Text Encoding" options.
During my investigation, I noticed certain patterns in how the corruption occurs. Specifically, I generated 350 text files using the following AutoHotkey v1 script:
Code: Select all
Loop 350 {
count := A_Index
Loop 6 {
if (A_Index == 1) {
paragraph := ""
Loop, % count
paragraph .= "あ"
output := paragraph
} else
output .= "`n`n" . paragraph
}
FileAppend, %output%, %count%.txt
}- Each text file consists of 6 paragraphs.
- Each paragraph consists of the Hiragana character "あ" (JIS code: 0x2422), with the character count increasing by one for each file.
- Paragraphs are separated by a blank line.
- The encoding is Japanese JIS (ISO-2022-JP) with LF line endings.
- Within the paragraphs, each character あ (JIS code: 0x2422) is incorrectly rendered as the string $". This results in long sequences of corruption such as $"$"$"$".
- The specific corruption string ←(B (likely related to the ESC (B escape sequence) appears at the very end of the text.
- Notably, this ←(B corruption only occurs when the text file ends without a trailing newline (e.g., ...あ[EOF]). If the file ends with a newline (e.g., ...あ[LF][EOF]), this specific trailing corruption does not occur.
- On pages 1 through 81, as well as on specific pages (pages 124 and 252), no character corruption occurs at all within the paragraphs.
Hoping that the above information will be of some help to you.
Thank you so much for your continued support.
Best regards,
rakunavi
- PDF-XChange Editor PRO Version: 10.8.4 build 409
- OS Version: Windows 11 Pro / Home 25H2 Build 26200.8037
- Adobe Acrobat X Standard Version 10.1.16
- PC Model: GMKtec Nucbox M7 Pro with HUION Kamvas Pro 19 / Lenovo IdeaPad C340-15IWL
[EDIT] Since a ticket has been officially issued for this request, I have changed the topic title as follows. 2026-3-27 6:07 JST (UTC+9)
- Previous Title: Character corruption when converting Japanese JIS (ISO-2022-JP) text files to PDF
- New Title: RT#7802: Bug: Convert JP text file to PDF errors