One of the easiest ways to introduce errors into Sindhi digital text is to choose a character because it looks right. Arabic-script alphabets contain many shapes that are visually close, especially at small sizes. Unicode keeps those characters separate when they have different identities.
Sindhi ڪ vs. Urdu/Persian ک
| Character | Code point | Unicode name | Important note |
|---|---|---|---|
| ڪ | U+06AA | ARABIC LETTER SWASH KAF | Distinct character used in Sindhi |
| ک | U+06A9 | ARABIC LETTER KEHEH | Used in Persian, Urdu and also appears in Sindhi text |
The Unicode chart explicitly notes that U+06AA represents a letter distinct from Arabic KAF in Sindhi. Verify the official Unicode annotation.
Yeh-family characters
Arabic-script yeh characters are another source of confusion because fonts can change the visual details between isolated, initial, medial, and final forms. When a text-processing problem is suspected, inspect the actual code point rather than comparing screenshots.
Dal-family characters
Sindhi has several extended Arabic letters in the U+0680–U+068F range. Unicode identifies characters such as ڊ (U+068A), ڌ (U+068C), ڍ (U+068D), and ڏ (U+068F) as Sindhi-related characters. The current keyboard includes several of these directly.
A reliable comparison method
- Copy the character.
- Look up the code point.
- Check the official Unicode name.
- Compare it with the visually similar alternative.
- Use the language-appropriate character for the word you are typing.
Use the Sindhi Unicode Character Reference for a broader table, and read Sindhi vs Urdu Script Differences for the linguistic context behind some of these distinctions.
Why visual comparison is unreliable
Fonts are free to draw characters in different styles. At small sizes, dots, hooks, tails, and other distinguishing features can become difficult to see. Two characters from different Unicode code points can therefore look almost identical in one font and clearly different in another.
This is why a language-aware Unicode reference is more reliable than a picture chart alone. When a spelling matters, identify the code point first and inspect the glyph second.
Useful comparisons to remember
| Look at | Code points | Why check |
|---|---|---|
| ڪ / ک | U+06AA / U+06A9 | Distinct characters used in related Arabic-script writing systems |
| ڊ / ڈ | U+068A / U+0688 | Different extended letters despite similar visual structure |
| ڌ / ڍ | U+068C / U+068D | Different Unicode characters with closely related shapes |
The official Unicode annotations are the best place to confirm the identity of a character. The full reference table is intended for repeated lookup, while this comparison page focuses on the mistakes that are easiest to make.
Primary References
This page uses the Unicode Standard as its technical reference where Unicode character identities, code points, bidirectional behavior, or Arabic-script annotations are discussed.
- Unicode Standard — Chapter 9: Arabic script
- Unicode 18.0 Arabic character names and annotations
- Unicode Standard Annex #9 — Bidirectional Algorithm
Reference links reviewed: September 2026