RE: Unicode normalization in XML 1.1

To: "'John Cowan'" <cowan@m...>,"'Lars Marius Garshol'" <larsga@g...>
Subject: RE: Unicode normalization in XML 1.1
From: "Michael Kay" <michael.h.kay@n...>
Date: Thu, 3 Apr 2003 16:50:06 +0100
Cc: <xml-dev@l...>
Importance: Normal
In-reply-to: <20030403132804.GI29046@c...>
Reply-to: <michael.h.kay@n...>

Play the video

> The point is that normalization is expensive, and it may be 
> too expensive to do at all in small systems.  Therefore, the 
> W3C's choice (expressed in the Character Model) is to have 
> senders normalize, and receivers check for normalization.  In 
> this way documents are normalized once at creation (or 
> publication) time, rather than every time a document is 
> received; this conserves net-wide cycles, since checking is 
> cheaper than normalizing.

While this policy makes sense, its translation into rules for software
components is unfortunately full of absurdities. The fact that the
character model [1] bans text processing software from doing
normalization [2] means that senders are going to have a tough job
meeting the requirement to normalize the text, because they won't be
able to find any text processing software that does the job for them.


[1] http://www.w3.org/TR/charmod/

[2] Section 4.4: "A text processing component .... must not normalize
suspect text".

Michael Kay

Follow-Ups:
- Re: Unicode normalization in XML 1.1
  - From: John Cowan <jcowan@r...>
- Re: Unicode normalization in XML 1.1
  - From: "Rick Jelliffe" <ricko@a...>

References:
- Re: Unicode normalization in XML 1.1
  - From: John Cowan <cowan@m...>

Prev by Date: RE: Design as, one hopes, not premature optimization
Next by Date: On the aparent importance of emoticons (was "Design as, one hopes, not premature optimization")
Previous by thread: Re: Unicode normalization in XML 1.1
Next by thread: Re: Unicode normalization in XML 1.1
Index(es):
- Date
- Thread

PURCHASE STYLUS STUDIO ONLINE TODAY!

Purchasing Stylus Studio from our online shop is Easy, Secure and Value Priced!

Download The World's Best XML IDE!

Accelerate XML development with our award-winning XML IDE - Download a free trial today!

Subscribe in XML format

RSS 2.0
Atom 0.3

Stylus Studio has published XML-DEV in RSS and ATOM formats, enabling users to easily subcribe to the list from their preferred news reader application.

Stylus Studio Sponsored Links are added links designed to provide related and additional information to the visitors of this website. they were not included by the author in the initial post. To view the content without the Sponsor Links please click here.

XML Editor - Download a 15 Day Free Trial Now >

See What's New in Stylus Studio >

Buy Stylus Studio - XML Editor - Now >