Unicode Overview:
- Coverage: 143,859 characters (as of Unicode 15.0)
- Languages: 159 modern and historic writing systems
- Includes: Every language, emoji, math symbols, ancient scripts, musical notation
- Backward compatible: First 128 codes match ASCII
- Goal: Every character in every language gets ONE unique code
Unicode Examples:
U+0041: A (Latin A) U+03A9: Ξ© (Greek Omega) U+4E2D: δΈ (Chinese "middle") U+0628: Ψ¨ (Arabic letter beh) U+1F600: π (Grinning face emoji) U+1F4A9: π© (Pile of poo emoji) U+00E9: Γ© (e with acute accent)
This key facts covers Unicode: The Universal Solution within Character Sets for GCSE Computer Science. Revise Character Sets in 3.3 Data Representation for GCSE Computer Science with 15 exam-style questions and 18 flashcards. This topic appears less often, but it can still be a useful differentiator on mixed-topic papers. It is section 5 of 11 in this topic. Use this key facts to connect the idea to the wider topic before moving on to questions and flashcards.
Practice questions for Character Sets
How many bits does standard ASCII use to represent each character?
Explain why using Unicode to store a text file produces a larger file than using ASCII to store the same text.