도구로 돌아가기
Read this page in English
자료 출처
글리팝의 기호 도감·이모지 도감·유니코드 탐색기에 나오는 글자 목록과 이름은
유니코드 컨소시엄이 공개한 공식 데이터 파일에서 가져왔습니다.
어느 파일에서 무엇을 가져왔는지, 어디까지가 공식 이름이고 어디부터가 저희가 붙인 이름인지 아래에 적었습니다.
1. 쓰고 있는 파일
-
UnicodeData.txt (유니코드 17.0) — 각 글자의 공식 영문 이름과 분류(문자·기호·문장부호 등).
기호 도감의 구역별 글자 수와 탐색기의 이름표가 모두 이 파일에서 나옵니다.
원본: unicode.org/Public/17.0.0/ucd/UnicodeData.txt
-
emoji-test.txt (이모지 17.0, UTS #51) — 이모지 목록과 갈래(표정·동물·음식 …), 그리고 공식 영문 이름.
이모지 도감의 목록이 이 파일입니다.
원본: unicode.org/Public/emoji/17.0/emoji-test.txt
두 파일은 만들 때 한 번만 읽습니다. 사이트를 쓰는 동안에는 유니코드 서버든 어디든
바깥으로 아무것도 불러오지 않습니다. 필요한 만큼만 미리 만들어 함께 담아 두었습니다.
2. 어떤 이름이 공식 이름인가
- 기호의 이름은 전부 유니코드 공식 영문 이름입니다. 저희가 지어낸 이름은 없습니다.
기호에는 한국어 이름표가 없어서, 이름으로 찾기는 영문 이름과 U+코드로만 됩니다.
- 이모지의 영문 이름도 유니코드 공식 이름입니다.
- 이모지의 한국어 이름은 글리팝이 직접 붙였습니다. 전부가 아니라 일부에만 있고,
한국어 이름이 없는 이모지는 영문 공식 이름으로 보여 줍니다. 화면에 그 수를 그대로 적어 둡니다.
- 구역 이름(「화살표」, 「점자」 …)과 갈래 이름(「표정과 감정」 …)의 한국어 표현도 글리팝이 붙인 것입니다.
영문 쪽은 유니코드가 쓰는 이름을 따릅니다.
3. 목록에서 빠지는 글자
구역 안이라도 유니코드가 배정하지 않은 빈자리, 화면에 보이지 않는 제어·서식 문자, 공백,
그리고 앞 글자에 붙어야만 쓰이는 결합 부호는 빼고 보여 줍니다. 눌러서 복사해도 아무것도 붙지 않기 때문입니다.
그래서 구역의 번호 범위보다 실제 글자 수가 적습니다. 화면에 적힌 수는 모두 이 파일에서 센 값입니다.
이모지는 정식 형태(fully-qualified)만 담습니다. 같은 그림을 조금 다르게 적은 형태는 중복이라 빼고,
피부색·머리색 변형은 각각 별개의 이모지로 그대로 담습니다.
4. 라이선스 표기
위 데이터 파일은 유니코드 라이선스 v3에 따라 씁니다. 전문은
여기에 그대로 담아 두었고,
원문은 unicode.org/license.txt 에 있습니다.
Unicode® 는 Unicode, Inc. 의 등록 상표입니다. 글리팝은 유니코드 컨소시엄과 관련이 없고,
유니코드의 승인을 받은 서비스도 아닙니다.
Back to the tool
이 페이지를 한국어로 읽기
Where the data comes from
The character lists and names in the symbol catalogue, the emoji catalogue and the Unicode
explorer all come from the data files the Unicode Consortium publishes. This page says which
file each part came from, and which names are Unicode's and which are ours.
1. The files
-
UnicodeData.txt (Unicode 17.0) - the formal English name and the general category of
every character. The per-block counts in the catalogue and the names in the explorer are
both read from this file.
Original: unicode.org/Public/17.0.0/ucd/UnicodeData.txt
-
emoji-test.txt (Emoji 17.0, UTS #51) - the emoji list, the groups they fall into, and
their official English names. The emoji catalogue is this file.
Original: unicode.org/Public/emoji/17.0/emoji-test.txt
Both files are read once, when the site is built. Nothing is fetched from unicode.org - or from
anywhere else - while you are using the site.
2. Which names are official
- Symbol names are the formal Unicode names, without exception. Nothing here is a name
we made up. The symbols have no Korean names, so searching by name works in English and by
U+ code only.
- Emoji English names are Unicode's own.
- Emoji Korean labels are written by this project. They exist for some entries, not
all; an emoji without one shows its official English name, and the screen says how many
have a label.
- Block names ("Arrows", "Braille patterns") and group names ("Smileys &
Emotion") follow Unicode in English; the Korean wording for them is ours.
3. What is left out
Inside a block, the unassigned holes are skipped, and so are the invisible things: controls,
format characters, spaces, and combining marks that only exist attached to the character before
them. Copying one of those gives you nothing, so they are not offered. That is why a block
holds fewer characters than its code point range suggests. Every count on the page is counted
from the data file.
For emoji, only the fully-qualified sequences are listed: the same picture written with fewer
variation selectors would be a duplicate. Skin tone and hair variants are separate sequences
and are kept as separate entries.
4. Licence
The data files are used under Unicode Licence v3. The full text is
included here, and the original is at
unicode.org/license.txt.
Unicode® is a registered trademark of Unicode, Inc. GlyPop is not affiliated with, or endorsed
by, the Unicode Consortium.