c++ - Is wchar_t needed for unicode support?

Question

Welcome To Ask or Share your Answers For Others

c++ - Is wchar_t needed for unicode support?

asked Oct 17, 2021 in Technique[技术] by 深蓝 (71.8m points)

Is the wchar_t type required for unicode support? If not then what's the point of this multibyte type? Why would you use wchar_t when you could accomplish the same thing with char?

See Question&Answers more detail:os

与恶龙缠斗过久,自身亦成为恶龙；凝视深渊过久,深渊将回以凝视…

328 views

1 Answer

深蓝 · Answer 1 · 2021-10-17T00:20:11+0000

No.

Technically, no. Unicode is a standard that defines code points and it does not require a particular encoding.

So, you could use unicode with the UTF-8 encoding and then everything would fit in a one or a short sequence of char objects and it would even still be null-terminated.

The problem with UTF-8 and UTF-16 is that s[i] is not necessarily a character any more, it might be just a piece of one, whereas with sufficiently wide characters you can preserve the abstraction that s[i] is a single character, tho it does not make strings fixed-length under various transformations.

32-bit integers are at least wide enough to solve the code point problem but they still don't handle corner cases, e.g., upcasing something can change the number of characters.

So it turns out that the x[i] problem is not completely solved even by char32_t, and those other encodings make poor file formats.

Your implied point, then, is quite valid: wchar_t is a failure, partly because Windows made it only 16 bits, and partly because it didn't solve every problem and was horribly incompatible with the byte stream abstraction.

Categories

c++ - Is wchar_t needed for unicode support?

Please log in or register to add a comment.

Please log in or register to answer this question.

1 Answer

No.

Please log in or register to add a comment.

Just Browsing Browsing

Most popular tags