Arabic PDF to Word: why the text comes out reversed and how to fix it
Last updated: 2026-09-25
Many people convert an Arabic PDF to Word and find reversed words, broken letters or missing spaces. The fault is not yours or your file’s; it comes from how the PDF format works.
How does a PDF store Arabic text?
A PDF stores the letter shapes drawn on the page (presentation forms: initial, medial, final) in display order from left to right, not the order you typed them. When text is extracted it appears reversed unless the tool rebuilds it.
What should the tool do?
- Convert the joined presentation forms back to normal Arabic letters.
- Reorder the words from right to left based on their position on the page.
- Keep digits and Latin words in their natural order.
- Group consecutive lines into paragraphs.
Scanned PDFs
If the file is a picture of a paper page, there is no text at all. You need optical character recognition (OCR), or you can settle for the page image inside a Word file.
The best practical approach
- Convert the file with a tool that supports Arabic, such as the PDF to Word tool.
- Keep the page image as a reference for the original look.
- Edit only the extracted text, then reformat it.
And if you copied text from a PDF and it looks broken, paste it into the Arabic text tools and turn on “Convert presentation forms”.
Summary
Reversed text is a known problem that can be solved by rebuilding the letters and the order. Choose a tool that handles Arabic properly, and keep the page image as a backup of the layout.