Baybayin Transliterator
Convert supplied Latin Tagalog spelling into Unicode Baybayin with a visible map of every source span. Choose modern final-consonant marks or explicit traditional omissions, review unsupported spellings and copy the actual Unicode text.
Up to 20,000 Unicode characters. Uppercase Latin letters map like lowercase. Accented clusters are checked in NFC and kept unchanged for review. Digits, punctuation and whitespace stay in place.
ᜀᜃᜓ ᜊᜌᜈ᜔ ᜋᜑᜎ᜔ ᜃᜒᜆ ᜉᜒᜎᜒᜉᜒᜈᜐ᜔ ᜅ᜔
Unicode code points: font-independent fallback
U+1700 U+1703 U+1713 U+0020 U+170A U+170C U+1708 U+1714 U+0020 U+170B U+1711 U+170E U+1714 U+0020 U+1703 U+1712 U+1706 U+0020 U+1709 U+1712 U+170E U+1712 U+1709 U+1712 U+1708 U+1710 U+1714 U+0020 U+1705 U+1714
Boxes or missing glyphs indicate system-font support. Copy and TXT export preserve the Unicode characters without requiring a particular font.
I/E and U/O share signs. This script output cannot recover the original Latin vowel distinction.
Original example shown. Transliterate to build a current syllable map.
- Mapped / omitted tokens
- 16
- Spelling review
- 0
- Omitted consonants
- 0
- Unicode characters
- 33 → 30
Spelling review and explicit edits
Positions count original Unicode characters from 1, including emoji, spaces and combining marks. Unsupported letters stay exactly as typed. Apply a respelling to replace that source span and regenerate the full map.
No unsupported spelling tokens need an edit. 5 mapped syllables use shared I/E or U/O signs; this ambiguity is inherent and cannot be reversed from the output alone.
Traditional omission report
No consonants were omitted. Modern modes retain final and cluster consonants with the selected cancellation mark.
Syllable and source-span map
Each row shows the exact original span. Spaces and line breaks are quoted so they remain visible in this table. NG counts as one consonant. Every mapped mark has its Unicode code point.
| Position | Original span | Output | Code points | Explanation |
|---|---|---|---|---|
| 1 | "a" | ᜀ | U+1700 | Consonant + A, or independent A. |
| 2–3 | "ko" | ᜃᜓ | U+1703 U+1713 | U and O share this sign; the Latin distinction is lost. |
| 4 | " " | " " | U+0020 | Punctuation, number, symbol or whitespace preserved exactly. |
| 5–6 | "ba" | ᜊ | U+170A | Consonant + A, or independent A. |
| 7–8 | "ya" | ᜌ | U+170C | Consonant + A, or independent A. |
| 9 | "n" | ᜈ᜔ | U+1708 U+1714 | Final or cluster consonant canceled with U+1714 VIRAMA. |
| 10 | " " | " " | U+0020 | Punctuation, number, symbol or whitespace preserved exactly. |
| 11–12 | "ma" | ᜋ | U+170B | Consonant + A, or independent A. |
| 13–14 | "ha" | ᜑ | U+1711 | Consonant + A, or independent A. |
| 15 | "l" | ᜎ᜔ | U+170E U+1714 | Final or cluster consonant canceled with U+1714 VIRAMA. |
| 16 | " " | " " | U+0020 | Punctuation, number, symbol or whitespace preserved exactly. |
| 17–18 | "ki" | ᜃᜒ | U+1703 U+1712 | I and E share this sign; the Latin distinction is lost. |
| 19–20 | "ta" | ᜆ | U+1706 | Consonant + A, or independent A. |
| 21 | " " | " " | U+0020 | Punctuation, number, symbol or whitespace preserved exactly. |
| 22–23 | "Pi" | ᜉᜒ | U+1709 U+1712 | I and E share this sign; the Latin distinction is lost. |
| 24–25 | "li" | ᜎᜒ | U+170E U+1712 | I and E share this sign; the Latin distinction is lost. |
| 26–27 | "pi" | ᜉᜒ | U+1709 U+1712 | I and E share this sign; the Latin distinction is lost. |
| 28–29 | "na" | ᜈ | U+1708 | Consonant + A, or independent A. |
| 30 | "s" | ᜐ᜔ | U+1710 U+1714 | Final or cluster consonant canceled with U+1714 VIRAMA. |
| 31 | " " | " " | U+0020 | Punctuation, number, symbol or whitespace preserved exactly. |
| 32–33 | "ng" | ᜅ᜔ | U+1705 U+1714 | NG is one consonant; no vowel is added. Final or cluster consonant canceled with U+1714 VIRAMA. |
Worked examples
- ᜊᜌᜈ᜔ · U+170A U+170C U+1708 U+1714
- ᜊᜌ · n at position 5 is omitted.
- ᜅ᜔ · U+1705 U+1714
Accuracy. Writing-system transliteration, not translation into Tagalog. I/E and U/O share signs; script output loses those Latin distinctions. Default modern mode retains final consonants with U1714; traditional mode visibly reports omitted consonants. No assertion of one universally correct historical spelling or reversible translation.
Common questions
- Is this a Baybayin translator for English?
- It is a writing-system transliterator for supplied Latin Tagalog orthography. It does not translate English vocabulary into Tagalog, infer pronunciation or invent a spelling for a name. Supply the spelling you intend, then inspect the syllable map and any review items. Foreign letters remain visible until you explicitly edit them.
- How are syllables and NG mapped?
- A consonant followed by a, e, i, o or u forms a mapped syllable. A has no added vowel mark; I/E share one mark and U/O share another. Independent vowels use their own signs. NG is one consonant, including the whole word ng: modern virama mode gives ᜅ᜔ without inventing a vowel or expanding it to another word.
- What is the difference between virama, pamudpod and traditional mode?
- Modern virama mode appends U+1714 to a consonant without a following vowel. The alternate modern mode uses U+1715 pamudpod. Traditional mode omits those consonants and lists each original span and position in its omission report. Bayan becomes ᜊᜌᜈ᜔ with virama and ᜊᜌ in traditional mode, with the final n at position 5 reported.
- Which R and D convention does the tool use?
- The default maps R to the explicitly labelled modern RA character U+170D. You can instead choose the convention where R shares DA, U+1707. D itself maps to DA in both modes. These are visible user choices, without claiming one universally correct historical spelling.
- How do I handle unsupported letters or accents?
- Letters such as f, v, z, q, c, x and j, along with other unsupported alphabetic or accented clusters, remain unchanged and are listed for review. Enter your own replacement spelling for a listed span and apply it to regenerate the output and positions. Ordinary letter case is ignored. Accented clusters are examined in NFC form but retain their original code points until you edit them; no accent or foreign letter is silently substituted.
- Can the output be converted back to the exact original spelling?
- No. I and E share signs, as do U and O, so those Latin distinctions are lost. The shared DA convention also merges R and D. Traditional omission removes consonants from the output. The source-span map preserves the supplied spelling for inspection, but the Baybayin string alone is not a reversible encoding of it.
- What happens to spaces, punctuation, digits and emoji?
- They remain in their original positions. The tool accepts up to 20,000 Unicode code points, counting an emoji outside the basic multilingual plane as one code point. Review positions count the original source, including spaces and combining marks. Empty or whitespace-only input is rejected, and incomplete surrogate characters must be corrected before a UTF-8 export can preserve the text.
- Why might I see boxes, and what do Copy and TXT export contain?
- Boxes usually mean the system font lacks the Baybayin glyph. The code-point fallback and per-span table identify the actual Unicode characters independently of fonts. Copy and UTF-8 TXT contain exactly the visible output string, including chosen cancellation marks, whitespace and unresolved spelling tokens. They do not contain images or a private font encoding. Processing is local, with no uploads or saved text.
Writing-system transliteration, not translation into Tagalog. I/E and U/O share signs; script output loses those Latin distinctions. Default modern mode retains final consonants with U1714; traditional mode visibly reports omitted consonants. No assertion of one universally correct historical spelling or reversible translation.