Description:
The Windows 4.1 alpha binary does not seem to be compiled with the support for UTF8 encoding.
C:\mysql\bin>mysqld --default-character-set=utf8
mysqld: File 'C:\mysql\share\charsets\utf8.xml' not found (Errcode: 2)
mysqld: Character set 'utf8' is not a compiled character set and is not specified in the 'C:\mysql\share\charsets\Index.xml' file
030411 15:10:28 Aborting
030411 15:10:28 mysqld: Shutdown Complete
The encodings listed by show character set list utf8 as a supported encoding:
mysql> SHOW CHARACTER SET;
+----------+-----------------------------+---------------------+--------+
| Charset | Description | Default collation | Maxlen |
+----------+-----------------------------+---------------------+--------+
| big5 | Big5 Traditional Chinese | big5 | 1 |
| dec8 | DEC West European | dec8_swedish_ci | 1 |
| cp850 | DOS West European | cp850_general_ci | 1 |
| hp8 | HP West European | hp8_english_ci | 1 |
| koi8r | KOI8-R Relcom Russian | koi8r_general_ci | 1 |
| latin1 | ISO 8859-1 West European | latin1_swedish_ci | 1 |
| latin2 | ISO 8859-2 Central European | latin2_general_ci | 1 |
| swe7 | 7bit Swedish | swe7_swedish_ci | 1 |
| ascii | US ASCII | ascii_general_ci | 1 |
| ujis | EUC-JP Japanese | ujis | 1 |
| sjis | Shift-JIS Japanese | sjis | 1 |
| cp1251 | Windows Cyrillic | cp1251_bulgarian_ci | 1 |
| hebrew | ISO 8859-8 Hebrew | hebrew | 1 |
| tis620 | TIS620 Thai | tis620 | 1 |
| euckr | EUC-KR Korean | euckr | 1 |
| koi8u | KOI8-U Ukrainian | koi8u_general_ci | 1 |
| gb2312 | GB2312 Simplified Chinese | gb2312 | 1 |
| greek | ISO 8859-7 Greek | greek | 1 |
| cp1250 | Windows Central European | cp1250_general_ci | 1 |
| gbk | GBK Simplified Chinese | gbk | 1 |
| latin5 | ISO 8859-9 Turkish | latin5_turkish_ci | 1 |
| armscii8 | ARMSCII-8 Armenian | armscii8_general_ci | 1 |
| utf8 | UTF-8 Unicode | utf8 | 1 |
| ucs2 | UCS-2 Unicode | ucs2 | 1 |
| cp866 | DOS Russian | cp866_general_ci | 1 |
| keybcs2 | DOS Kamenicky Czech-Slovak | keybcs2 | 1 |
| macce | Mac Central European | macce | 1 |
| macroman | Mac West European | macroman | 1 |
| cp852 | DOS Central European | cp852_general_ci | 1 |
| latin7 | ISO 8859-13 Baltic | latin7_general_ci | 1 |
| cp1256 | Windows Arabic | cp1256_general_ci | 1 |
| cp1257 | Windows Baltic | cp1257_ci_ai | 1 |
| binary | Binary pseudo charset | binary | 1 |
+----------+-----------------------------+---------------------+--------+
33 rows in set (0.00 sec)
The share\charsets directory does not have a utf8.xml file. The index.xml fiels does have an entry for the name utf8 though.
If you start the server normally and then try to alter a database to use the utf8 charset, then you get the following error:
mysql> ALTER DATABASE temp DEFAULT CHARACTER SET utf8;
ERROR 1115: Unknown character set: 'utf8'
How to repeat: