Unicode to Non Unicode Problems — Causes and Fixes

The 5 conversion failures Telugu users hit most, what causes each one, and how to fix it.

Paste your text Diagnosis
Try:

Paste text above and the checker will identify its encoding.

100% Free
No Signup Required
Secure — Runs in Browser

There are 5 common Unicode to non Unicode conversion problems: Telugu text showing boxes, garbled output, the Anu font not displaying, wrong conjuncts or matras, and encoding mismatch after copy-paste. Each one is a different Telugu font conversion error with a different fix.

All 5 problems trace back to one of 3 root causes: the wrong Anu version was selected, the required font is not installed, or a copy-paste stripped the encoding. Identifying which of the 3 applies solves the problem faster than trying fixes at random. Most reports of Unicode to non Unicode not working reduce to the first cause.

This page covers the cause and fix for each of the 5 problems, how to verify a conversion is correct, and how to avoid these non Unicode text issues in future conversions.

Why Unicode to Non Unicode Conversion Fails

Conversion fails for 3 reasons: a version mismatch, a missing font, or a mixed-encoding paste. Every symptom on this page reduces to one of them.

Version mismatch

Anu 6 and Anu 7 assign different bytes to the same characters. Decoding or encoding with the wrong version returns wrong Telugu.

Missing font

Anu Script needs its font installed to render. Unicode Telugu needs any Telugu font. Neither renders in a font that lacks the glyphs.

Mixed-encoding paste

Pasting converted text beside unconverted text leaves one paragraph needing two fonts. No single font renders both runs.

Problem 1 — Telugu Font Showing Boxes or Question Marks

?

Cause

Telugu font showing symbols, telugu font showing question marks and telugu text showing boxes are three reports of two different failures. Boxes mean the display font has no glyph for those code points — the text is intact and the font cannot draw it. Question marks mean the text passed through a container that could not hold Telugu at all, such as a VARCHAR database column or a file saved as ANSI, and the characters were replaced at write time.

Fix

Apply a Telugu font that covers the characters: Pothana2000, Gautami or Noto Sans Telugu for Unicode text. Set the document or page charset to UTF-8, if boxes survive the font change. Recover question-marked text from the original source, because the characters were destroyed rather than mis-displayed. Telugu characters not showing at all points to the same write-time loss.

Problem 2 — Converted Telugu Text Is Garbled or Wrong

?

Cause

The Anu version selected during conversion does not match the font applied to the output. Anu 6 and Anu 7 assign different bytes to the same Telugu characters, and every gunintam form differs between them. Output encoded for one version rendered in the other produces unrelated letters and symbols.

Fix

Open the source file and read the font name in the character panel, because the name carries the version number. Convert again with that version selected. Try Anu 7 first, if the version is unknown, because most current files use it.

Problem 3 — Anu Font Not Displaying After Conversion

?

Cause

Anu Script output looks like Latin letters until an Anu font is applied to it. A browser has no Anu font installed, so correct output still displays as accented Latin characters there. The conversion is not at fault.

Fix

Paste the output into PageMaker, CorelDraw or Word and apply an Anu font to that text. Install Anu Script Manager (ASM), or the Anu font files on their own, if the font is missing from the machine.

Problem 4 — Telugu Conjuncts or Matras Are Wrong

?

Cause

Conjuncts break first when the Anu version is wrong, because that is where the two mappings diverge most. Every gunintam form — a consonant joined to a vowel sign — encodes differently in Anu 6 and Anu 7, so Anu conjuncts wrong is the first symptom of a version mismatch. Rare conjuncts falling outside the mapping table pass through unchanged rather than converting.

Fix

Switch the Anu version and convert again. Type the conjunct directly in your layout application for the few characters that fall outside the table. Check క్ష and జ్ఞ first, because they expose a version mismatch immediately.

Problem 5 — Encoding Mismatch After Copy-Paste

?

Cause

An encoding mismatch appears when one paragraph carries both Unicode text and non Unicode text. An encoding mismatch Telugu users hit most often comes from pasting a converted phrase into an unconverted paragraph: two runs need two different fonts, and no single font renders both correctly.

Fix

Convert the whole block rather than selected words. Check which font each text run carries, and apply one encoding across the paragraph. Keep an untouched copy of the original before replacing text in a layout.

How to Verify Your Conversion Is Correct

To verify a conversion, run one short word through both directions and compare the result against the original. There are 4 checks.

1

Round-trip one word

Convert a single word, then convert the output back. Matching text confirms the mapping and the version.

2

Test a conjunct

Use క్ష or జ్ఞ. Conjuncts expose a version mismatch immediately, where plain consonants do not.

3

Check the character count

Anu output is usually longer than the Telugu input, because one character can map to several bytes.

4

Apply the font

Paste into the layout application with the matching Anu font before converting the full document.

Test before you commit. Convert one word and check it renders, then convert the full document. Recovering a wrongly converted 200-page layout costs far more than the 30 seconds this check takes.

Conclusion

The 5 common Unicode to non Unicode problems are boxes or question marks, garbled output, the Anu font not displaying, wrong conjuncts or matras, and encoding mismatch after copy-paste. Each traces back to a version mismatch, a missing font, or a mixed-encoding paste.

Check the Anu version first, because it causes the two most-reported failures. Read Anu 6 vs Anu 7 to identify which version a file uses, and run conversions on the Unicode to Non Unicode Converter.

Frequently Asked Questions

Answers to 8 common questions about conversion problems.

Boxes mean the display font has no glyph at those code points. The text itself is intact. Apply a Telugu font that covers the characters: Pothana2000 or Gautami for Unicode output, or the matching Anu font for Anu Script output.

Garbled Telugu almost always means the Anu version was wrong. Anu 6 and Anu 7 assign different bytes to the same characters, so decoding with the wrong version returns unrelated symbols. Convert again with the other version selected.

The Anu font is either not installed on the machine, or it is installed but not applied to the text. Anu Script output looks like Latin letters until an Anu font is applied to it.

An encoding mismatch happens when one document holds both Unicode text and non Unicode text. Pasting a converted phrase into an unconverted paragraph produces half-readable output, because each run needs a different font.

To fix wrong conjuncts, switch the Anu version and convert again. Conjuncts such as క్ష and జ్ఞ are where Anu 6 and Anu 7 diverge most, so they break first when the version is wrong.

Question marks mean the text passed through a non Unicode container that could not hold Telugu. A VARCHAR database column or an ANSI-saved file replaces unsupported characters with question marks, and the original characters cannot be recovered.

To verify a conversion, convert one short word, paste the output back through the reverse direction, and compare it against the original. Matching text confirms the version and the mapping are right.

Selecting the wrong Anu version is the most common error. It accounts for garbled output and wrong conjuncts, which together are the two problems users report most.

Convert Telugu Without the Guesswork

Both Anu versions, both directions, with a round-trip check built in.

Check Your Text →