On Jan 21, 2008, at 3:43 PM, Dmitry Potapov wrote:
On Mon, Jan 21, 2008 at 11:59:24AM -0500, Kevin Ballard wrote:
quoted
No, it's a question of hashing algorithm. And it's one that's fairly
easily solved simply by picking a specific nonambiguous UTF-8
encoding
before hashing.
UTF-8 is a *single* encoding, and it maps every Unicode character to
a unique binary representation. So, it is completely nonambiguous.
In this case, encoding refers to normalization form, as other people
have used it in the conversation besides me.
I suggest you stop trying to find inconsequential stuff to argue
about, especially when a tiny bit of critical thinking would reveal
the answer.
-Kevin Ballard
--
Kevin Ballard
http://kevin.sb.org
kevin@sb.org
http://www.tildesoft.com