如何在不手动指定编码的情况下获得C#中字符串的一致字节表示？

如何在.NET（C#）中将字符串转换为字节[]而不手动指定特定编码？

我要加密字符串。我可以在不进行转换的情况下对其进行加密，但我仍然想知道为什么编码会在这里发挥作用。

此外，为什么还要考虑编码？我不能简单地获取字符串存储的字节数吗？为什么依赖字符编码？

当前回答

由于以下事实，字符串可以通过几种不同的方式转换为字节数组：.NET支持Unicode，Unicode标准化了几种称为UTF的不同编码。它们具有不同长度的字节表示，但在这个意义上是等价的，即当字符串被编码时，它可以被编码回字符串，但如果字符串用一个UTF编码，并且在不同UTF的假设下解码，如果可能会出错。

此外，.NET支持非Unicode编码，但它们在一般情况下无效（只有在实际字符串（如ASCII）中使用有限的Unicode代码点子集时才有效）。在内部，.NET支持UTF-16，但对于流表示，通常使用UTF-8。它也是互联网的事实标准。

毫不奇怪，System.Text.Encoding类是一个抽象类，它支持将字符串序列化为字节数组和反序列化；它的派生类支持具体编码：ASCIIEncoding和四个UTF（System.Text.UnicodeEncoding支持UTF-16）

参考此链接。

对于使用System.Text.Encoding.GetBytes对字节数组进行序列化。对于反向操作，使用System.Text.Encoding.GGetChars。此函数返回字符数组，因此要获取字符串，请使用字符串构造函数System.string（char[]）。请参阅本页。

例子：

string myString = //... some string

System.Text.Encoding encoding = System.Text.Encoding.UTF8; //or some other, but prefer some UTF is Unicode is used
byte[] bytes = encoding.GetBytes(myString);

//next lines are written in response to a follow-up questions:

myString = new string(encoding.GetChars(bytes));
byte[] bytes = encoding.GetBytes(myString);
myString = new string(encoding.GetChars(bytes));
byte[] bytes = encoding.GetBytes(myString);

//how many times shall I repeat it to show there is a round-trip? :-)

2014-06-11 11:29:06

其他回答

byte[] strToByteArray(string str)
{
    System.Text.ASCIIEncoding enc = new System.Text.ASCIIEncoding();
    return enc.GetBytes(str);
}

2009-01-23 13:43:18

对于串行通信项目，我必须将字符串转换为字节数组-我必须处理8位字符，而且我无法找到使用框架转换器的方法，这样既不会添加两个字节条目，也不会错误地转换具有第八位集的字节。所以我做了以下工作：

string message = "This is a message.";
byte[] bytes = new byte[message.Length];
for (int i = 0; i < message.Length; i++)
    bytes[i] = (byte)message[i];

2016-01-21 17:19:03

这取决于字符串的编码（ASCII、UTF-8…）。

例如：

byte[] b1 = System.Text.Encoding.UTF8.GetBytes (myString);
byte[] b2 = System.Text.Encoding.ASCII.GetBytes (myString);

编码重要的一个小例子：

string pi = "\u03a0";
byte[] ascii = System.Text.Encoding.ASCII.GetBytes (pi);
byte[] utf8 = System.Text.Encoding.UTF8.GetBytes (pi);

Console.WriteLine (ascii.Length); //Will print 1
Console.WriteLine (utf8.Length); //Will print 2
Console.WriteLine (System.Text.Encoding.ASCII.GetString (ascii)); //Will print '?'

ASCII根本无法处理特殊字符。

在内部，.NET框架使用UTF-16表示字符串，因此，如果您只想获得.NET使用的确切字节，请使用System.Text.Encoding.Unicode.GetBytes（…）。

有关详细信息，请参阅.NET Framework（MSDN）中的字符编码。

2009-01-23 13:43:51

可以使用以下代码将字符串转换为.NET中的字节数组

string s_unicode = "abcéabc";
byte[] utf8Bytes = System.Text.Encoding.UTF8.GetBytes(s_unicode);

2013-09-02 11:21:11

Use:

    string text = "string";
    byte[] array = System.Text.Encoding.UTF8.GetBytes(text);

结果是：

[0] = 115
[1] = 116
[2] = 114
[3] = 105
[4] = 110
[5] = 103

2013-10-22 12:55:59

如何在不手动指定编码的情况下获得C#中字符串的一致字节表示？

推荐文章

最新文章

标签