How can I find the character code of a special character in my text editor?

Question

When pasting text from outside sources into a plain-text editor (e.g. TextMate or Sublime Text 2) a common problem is that special characters are often pasted in as well. Some of these characters render fine, but depending on the source, some might not display correctly (usually showing up as a question mark with a box around it).

So this is actually 2 questions:

Given a special character (e.g., ’ or ♥) can I determine the UTF-8 character code used to display that character from inside my text editor, and/or convert those characters to their character codes?
For those "extra-special" characters that come in as garbage, is there any way to figure out what encoding was used to display that character in the source text, and can those characters somehow be converted to UTF-8?

Rob Napier · Accepted Answer

My favorite site for looking up characters is fileformat.info. They have a great Unicode character search that includes a lot of useful information about each character and its various encodings.

If you see the question mark with a box, that means you pasted something that can't be interpreted, often because it's not legal UTF-8 (not every byte sequence is legal UTF-8). One possibility is that it's UTF-16 with an endian mode that your editor isn't expecting. If you can get the full original source into a file, the file command is often the best tool for determining the encoding.

ndp · Answer

At &what I built a tool to focus on searching for characters. It indexes all the Unicode and HTML entity tables, but also supplements with hacker dictionaries and a database of keywords I've collected, so you can search for words like heart, quot, weather, umlaut, hash, cloverleaf and get what you want. By focusing on search, it avoids having to hunt around the Unicode pages, which can be frustrating. Give it a try.

How can I find the character code of a special character in my text editor?

Tags:

text

character-encoding

sublimetext2

utf-8

textmate

Aaron Fowler

Video Answer

2 Answers

Rob Napier

ndp

Recent Activity

Donate For Us

How can I find the character code of a special character in my text editor?

Tags:

text

character-encoding

sublimetext2

utf-8

textmate

Aaron Fowler

Video Answer

2 Answers

Rob Napier

ndp

Related questions

Recent Activity

Donate For Us