Arabic PDF to Word: why the text comes out reversed and how to fix it

Last updated: 2026-09-25

Many people convert an Arabic PDF to Word and find reversed words, broken letters or missing spaces. The fault is not yours or your file’s; it comes from how the PDF format works.

How does a PDF store Arabic text?

A PDF stores the letter shapes drawn on the page (presentation forms: initial, medial, final) in display order from left to right, not the order you typed them. When text is extracted it appears reversed unless the tool rebuilds it.

What should the tool do?

  • Convert the joined presentation forms back to normal Arabic letters.
  • Reorder the words from right to left based on their position on the page.
  • Keep digits and Latin words in their natural order.
  • Group consecutive lines into paragraphs.

Scanned PDFs

If the file is a picture of a paper page, there is no text at all. You need optical character recognition (OCR), or you can settle for the page image inside a Word file.

The best practical approach

  1. Convert the file with a tool that supports Arabic, such as the PDF to Word tool.
  2. Keep the page image as a reference for the original look.
  3. Edit only the extracted text, then reformat it.

And if you copied text from a PDF and it looks broken, paste it into the Arabic text tools and turn on “Convert presentation forms”.

Note: no tool can rebuild a complex table or layout from a PDF perfectly. That is why we put the page image next to the text.

Summary

Reversed text is a known problem that can be solved by rebuilding the letters and the order. Choose a tool that handles Arabic properly, and keep the page image as a backup of the layout.