Unicode Data & Sources
This page documents where our emoji and symbol data comes from and how it is prepared. SymbolTap is an independent project and is not an official Unicode Consortium service.
Sources
- Emoji characters, names and code points follow the public Unicode Standard and CLDR naming conventions.
- Symbol names and code points follow the Unicode character database.
- Keywords and short names are based on common public conventions and adjusted for search.
- Meaning descriptions on detail pages are written by us and are not copied from other sites.
Version & scope
- Target emoji data version: Unicode 15.1.
- Last reviewed: 2025-07-01.
- Curated entries currently published: 347 emojis plus a growing set of text symbols.
How the data is processed
- Code points are derived directly from each character to avoid transcription errors.
- Slugs and search keywords are normalized (lower-cased, de-duplicated).
- Skin-tone variants are generated for supported people and hand emojis.
- Duplicate entries are avoided, and each item is validated by automated tests.
Licensing
The Unicode Character Database is made available by the Unicode Consortium under its own terms. Emoji artwork is provided by your device’s operating system; we do not bundle or redistribute any vendor’s emoji images. See our Licenses page for details.