Wednesday, August 26, 2009

[Unicode Announcement] Last Call for Unicode 5.2 Data

The data files in the Unicode Character Database for Unicode 5.2 have
been revised to include all of the authorized changes from the last UTC
meeting. If you use any of the Unicode data in your implementations,
please update a test version of your implementation to use those files
and run your tests. If there are any showstopper bugs, please report
them (using http://www.unicode.org/reporting.html) as soon as possible.

From this point, the only adjustments that will be made to the data
will be on the basis of showstopper bugs, including bugs uncovered in
the process of updating the Unicode Collation data files for UCA 5.2.

----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

[Unicode Announcement] 33rd Internationalization & Unicode Conference - Keynote Speaker Announcement

33rd Internationalization & Unicode Conference
Features Sessions on Security, Open Source,
Social Networking and Cloud Computing

The Unicode® Consortium announces that Nicholas Ostler, Chairman, Foundation for Endangered Languages, will keynote the 33rd Internationalization & Unicode® Conference (IUC). The conference, sponsored by Gold Sponsors Adobe, Inc. and WinSoft, will take place in San Jose, Calif., USA; October 14-16, 2009. For more information and to register please visit www.unicodeconference.org/keynote-e.

Mr. Ostler will present "The Alphabetic Principle and its Enemies."

The alphabetic principle for writing seems brilliantly simple, and its implementation, often subverting other options, has often caused explosive growths in literacy, with important historical consequences for cultural survival. Its great advantages are economy of effort in the learner, and ready application to new languages. However, it has drawbacks as to speed for the initiated user, and also (by being essentially mechanical and phonetic) in representing many of the cultural overtones which people like their written language to have. There is, too, a certain resistance to the role of art in writing. But as alphabetic traditions age, becoming less purely alphabetic, these disadvantages can be reduced. New structures may emerge, meaningful patterns that leave alphabets far behind. Alphabetic scripts have more recently revealed new aspects, defining a convenient order to index anything, inspiring the phonemic principle of structural linguistics, and later mapping more easily!
than other systems onto digital systems, and hence a whole new set of functions for written language. But the alphabet remains a rather arbitrary means of representing meanings, since its icons are parasitic on the particular sounds of particular words in particular languages, a long way from thoughts.

About the Keynote Presenter
---------------------------
Nicholas Ostler holds an MA in classics, philosophy and economics from Oxford, and a PhD in linguistics from MIT. His first job was teaching in Japan, later consulting on machine translation for Fujitsu. Returning to England, he worked in IT research during the 1980s and '90s, especially with the UK government, and the European Union. He has been Chairman of the Foundation for Endangered Languages (www.ogmios.org) since its inception in 1996. He also edited its newsletter Ogmios until 2006. Within descriptive linguistics, his main research field has been the grammar of the (extinct) Chibcha language of Colombia. He has served on the board of the British National Corpus, the LSA's Committee for Endangered Languages, and on the editorial board of the International Journal of American Linguistics. As a writer, his book "Empires of the Word: a language history of the world" (HarperCollins, 2005) traced the histories of the large literate languages, from Sumerian to English, cons!
idering the factors that make for large-scale expansion. Later, "Ad Infinitum: a biography of Latin" (Walker & Co., 2007) considered the attitudes that have accompanied the Latin language throughout its 2,500 year recorded history. He is now at work on a book about the prospects of English as a global lingua franca, in the light of the competition, past and present. This is due for publication in 2010.

About the Internationalization & Unicode Conference
---------------------------------------------------
The Internationalization & Unicode Conference is the premier technical conference for both software and Web internationalization. Unicode and internationalization experts, implementers, clients and vendors are invited to attend this unique conference. The program committee has created an exciting program full of new and cutting-edge topics that is relevant and engaging for the internationalization community. The three-day conference will feature a full day of tutorials followed by two days of presentations, panels and discussions. There will also be technology exhibits and demonstrations. The interactive format makes the Internationalization & Unicode Conference a great place to meet and exchange ideas with leading experts, find out about the needs of potential clients, or get information about new and existing Unicode and internationalization-enabled products.
The 33rd Internationalization & Unicode Conference is sponsored by Gold Sponsors Adobe, Inc. and WinSoft; Media Sponsors LISA Globalization Insider and MultiLingual Computing Inc. and Organizational Sponsor Localization Industry Standards Association (LISA).

 
The early-bird registration deadline is September 4, 2009; the hotel registration deadline is September 23, 2009. For full conference details and to register, please click here. 
Sponsorships and exhibit space are available; for more information on sponsoring contact Ken Berk at kenberk@omg.org, +1-781-444 0404. For exhibiting questions email event_marketing@omg.org. For all other questions email info@unicodeconference.org.

About The Unicode Consortium
----------------------------
The Unicode Consortium is a non-profit organization founded to develop, extend and promote use of the Unicode Standard and related globalization standards.

The membership of the consortium represents a broad spectrum of corporations and organizations in the computer and information processing industry. Members are: Adobe Systems, Apple, DENIC eG, Google, Government of India, Government of Tamil Nadu, IBM, Microsoft, Monotype Imaging, Oracle, The Society for Natural Language Technology Research, Sun Microsystems, Sybase, The University of California at Berkeley, Yahoo!, plus well over a hundred Associate, Liaison, and Individual members.

For more information, please contact the Unicode Consortium www.unicode.org/contacts.html.

About the Event Producer
------------------------
OMGT is the Event Producer for the Internationalization & Unicode Conferences. OMG is an open membership, not-for-profit consortium that produces and maintains computer industry specifications for interoperable enterprise applications. Our specifications include MDA®, UML®, CORBA®, MOFT, XMI® and CWMT. OMG's specifications are all available for download by everyone without charge.

For more information about OMG, visit us online at www.omg.org.

If you would prefer not to receive messages from the OMG, or have address corrections, please reply to this email message, requesting Unsubscribe or describing your address corrections in the body of the text. Please leave subject line intact.

----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

Tuesday, July 28, 2009

[Unicode Announcement] Unicode 5.2 Beta - Chapters 1-5 Available

The ongoing beta review for Unicode 5.2 has been supplemented
today with the availability of drafts for the first part of the
consolidated text of the Unicode Standard, Version 5.2.

The landing page for Unicode 5.2 summarizes the major
new additions and changes for Version 5.2:

http://www.unicode.org/versions/Unicode5.2.0/

Links to pdf versions of Chapters 1 through 5 of the standard
are available on that page.

We would like to remind folks that the period for beta review
of the Version 5.2 data files and the Unicode Standard Annexes
is rapidly drawing to a close. The meeting of the Unicode
Technical Committee in August will be making the final decisions
on any reported problems in the data or the annexes, so now
is the time to check the posted data files and documents.
See the beta review page for details:

http://www.unicode.org/versions/beta.html

==================================================

If you have comments for official UTC consideration, please post them by
submitting your comments through our feedback & reporting page:

http://www.unicode.org/reporting.html

If you wish to discuss issues on the Unicode mail list, then please
use the following link to subscribe (if necessary). Please be aware
that discussion comments on the Unicode mail list are not automatically
recorded as input to the UTC. You must use the reporting link above
to generate comments for UTC consideration.

http://www.unicode.org/consortium/distlist.html

----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

Wednesday, July 15, 2009

[Unicode Announcement] 33rd Internationalization & Unicode Conference - Program Online

33rd Internationalization & Unicode Conference
Features Sessions on Security, Open Source, Social Networking and Cloud Computing

The Unicode(r) Consortium announces the program for the 33rd Internationalization & Unicode(r) Conference (IUC). The conference, sponsored by Gold Sponsor Adobe, Inc., will take place in San Jose, Calif., USA; October 14-16, 2009. The conference program is available online here.

The program committee has created an exciting program full of new and cutting-edge topics that is relevant and engaging for the internationalization community. The three-day conference will feature a full day of tutorials followed by two days of presentations, panels and discussions. There will also be technology exhibits and demonstrations.

Highlights of the Conference:
============================
Tutorials in Three Tracks:
-------------------------
An Introduction to Writing Systems & Unicode
Internationalization: An Introduction
Building a Custom Keyboard Layout for the Mac with Ukulele and XML
Arabic Script: Structure, Geographic and Regional Classification
Unicode - a Grand Tour
Web Internationalization - Standards and Best Practices
Building Multilingual Websites in Joomla [Drupal]
Creating XHTML/HTML Pages with Right-to-Left Scripts
Free Software stack for Unicode Text Rendering
Presenters come from such organizations as DecoType, Amazon, Penn State, Red Hat/GNOME, W3C, XenCraft, and Yahoo! Inc.

Sessions in Three Tracks:
------------------------
* Session tracks are categorized by Programming Languages, Fonts and Typography, Unicode News, and I18n Standards News on Thursday morning; with Open Source Libraries, Assuring Quality, and Scripts tracks in the afternoon. On Friday, track topics include Development Platforms, Mobile Programming, Internationalization in Practice, and Leveraging CLDR in the morning; and Translation Services API, Bidirectional Text, and Case Studies in the afternoon.

* The following is just a small sample of some of the cutting-edge presentations that will be given at IUC 33. For the full program, visit the IUC 33 Web site.

- Internationalization for JavaScript applications
- Emoji in Unicode: Cell Phones Meet the Internet
- Banking in the Cloud: Challenges of Internationalizing Banking Software
- Twanguages of the World: a Language Census of Twitter
- HarfBuzz, the Free and Open OpenType Shaping Engine


* Session Presenters come from such organizations as Adobe, Inc.; Amazon; Apple, Inc.; Casaba Security; DataDirect Technologies; DecoType; UC Berkeley; Google Inc.; HighTech Passport; IBM; Intel Corporation; Microsoft; Monotype Imaging; University of Michigan; XenCraft; Yahoo! Inc.; and Yale University
The Internationalization & Unicode Conference is the premier technical conference for both software and Web internationalization. Unicode and internationalization experts, implementers, clients and vendors are invited to attend this unique conference. The interactive format makes the Internationalization & Unicode Conference a great place to meet and exchange ideas with leading experts, find out about the needs of potential clients, or get information about new and existing Unicode and internationalization-enabled products.

Gold sponsor: ADOBE - Media Sponsor: Multilingual

The early-bird registration deadline is September 4, 2009; the hotel registration deadline is September 23, 2009. For full conference details and to register, please click here. Sponsorships and exhibit space are available; for more information on sponsoring contact Ken Berk at kenberk@omg.org, +1-781-444 0404. For exhibiting questions email event_marketing@omg.org. For all other questions email info@unicodeconference.org.


--------------------------------------------------------------------------------

About The Unicode Consortium
The Unicode Consortium is a non-profit organization founded to develop, extend and promote use of the Unicode Standard and related globalization standards.
The membership of the consortium represents a broad spectrum of corporations and organizations in the computer and information processing industry. Members are: Adobe Systems, Apple, DENIC eG, Google, Government of India, Government of Tamil Nadu, IBM, Microsoft, Monotype Imaging, Oracle, SAP, The Society for Natural Language Technology Research, Sun Microsystems, Sybase, The University of California at Berkeley, Yahoo!, plus well over a hundred Associate, Liaison, and Individual members.
For more information, please contact the Unicode Consortium www.unicode.org/contacts.html.

About the Event Producer
OMG(tm) is the Event Producer for the Internationalization & Unicode Conferences. OMG is an open membership, not-for-profit consortium that produces and maintains computer industry specifications for interoperable enterprise applications. Our specifications include MDA(r), UML(r), CORBA(r), MOF(tm), XMI(r) and CWM(tm). OMG's specifications are all available for download by everyone without charge.
For more information about OMG, visit us online at www.omg.org.

If you would prefer not to receive messages from the OMG, or have address corrections, please reply to this email message, requesting Unsubscribe or describing your address corrections in the body of the text. Please leave subject line intact.

----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

Wednesday, July 8, 2009

[Unicode Announcement] New Public Review #149: UTS #22; and other PRI updates

The Unicode Technical Committee has posted a new issue for public review
and comment. Details are on the following web page:

http://www.unicode.org/review/

Review period for the new item closes on August 3, 2009.

Please see the page for links to discussion and relevant documents.
Briefly, the new issue is:


PRI #149
Proposed Update UTS #22: Unicode Character Mapping Markup Language

This proposed update includes editorial fixes and clarifications based
on community feedback. There is a small change in the DTD from version
three to this proposed version five (a new default attribute value). See
the Modification History and the highlighted changes for details.


There are also two updates to open Public Review Issues:


PRI #136
Proposed Update UAX #14: Unicode Line Breaking Algorithm

The text of UAX #14 has been revised throughout, with both substantive
and editorial changes. A new Line_Break class CP has been added, and the
rule LB30 has been reintroduced, to address an edge case involving
breaks around parenthesized letters. More new Southeast Asian scripts
and characters have been added to the Line_Break class SA. The lists of
characters representing each Line_Break class are now exemplary, rather
than exhaustive in the text. Please review the new text carefully.


PRI #134
Proposed Update UAX #9: Unicode Bidirectional Algorithm

The latest revision also includes a new conformance test file, which
implementers should carefully review. See BidiTest.txt in the data files
directory: http://www.unicode.org/Public/5.2.0/ucd/


The closing dates remain the same.

If you have comments for official UTC consideration, please post them by
submitting your comments through our feedback & reporting page:

http://www.unicode.org/reporting.html

If you wish to discuss issues on the Unicode mail list, then please use
the following link to subscribe (if necessary). Please be aware that
discussion comments on the Unicode mail list are not automatically
recorded as input to the UTC. You must use the reporting link above to
generate comments for UTC consideration.

http://www.unicode.org/consortium/distlist.html

----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

Wednesday, July 1, 2009

[Unicode Announcement] Draft code charts for Unicode 5.2 beta review

Draft code charts are now available for the Unicode 5.2 beta review.
Please check the code charts carefully to verify correctness of the new
characters added to Unicode 5.2 and to ensure that there are no
regressions for previously encoded characters. The draft code charts
are located in:

http://www.unicode.org/Public/5.2.0/charts/ or
ftp://www.unicode.org/Public/5.2.0/charts/

The Unicode Consortium appreciates the help provided by its many volunteers
who help in ensuring the best possible quality for the published code charts
for the Unicode Standard.

For further information about the beta, please see the beta page
http://www.unicode.org/versions/beta.html and the associated
Public Review Issues page: http://www.unicode.org/review/#148


----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

Monday, June 29, 2009

[Unicode Announcement] Unicode Releases Common Locale Data Repository, Version 1.7.1

Mountain View, CA, June 29, 2009 - The Unicode® Consortium announced
today the release of a new version of the Unicode Common Locale Data
Repository (Unicode CLDR 1.7.1), providing key building blocks for
software to support the world's languages. Unicode CLDR is by far the
largest and most extensive standard repository of locale data. This data
is used by a wide spectrum of companies for their software
internationalization and localization: adapting software to the
conventions of different languages for such common software tasks as
formatting of dates, times, time zones, numbers, and currency values;
sorting text; choosing languages or countries by name; transliterating
different alphabets; and many others.

------------------------------------------------------------------------

CLDR 1.7.1 is an update release, with no new translations. The main
changes are fixes for numbering systems and currencies, but a number of
other bugs were fixed. See the CLDR 1.7.1 Release Note
<http://sites.google.com/site/cldr/index/downloads/cldr-1-7-1> for a
full list of changes. There were no changes in the LDML specification.

------------------------------------------------------------------------

Unicode CLDR 1.7 is part of the Unicode locale data project, together
with the Unicode Locale Data Markup Language (LDML:
http://unicode.org/reports/tr35/). LDML is an XML format used for
general interchange of locale data, such as in Microsoft's .NET. For web
pages with different views of CLDR data, see
http://unicode.org/cldr/charts.html. For more information about the
Unicode CLDR project (including charts) see http://cldr.unicode.org. The
latest features of CLDR will also be showcased at the 33rd
Internationalization and Unicode Conference (IUC) on October 14-16, 2009
in San Jose, CA — see http://unicodeconference.org/
<http://www.unicodeconference.org/>.

About the Unicode Consortium

The Unicode Consortium is a non-profit organization founded to develop,
extend and promote use of the Unicode Standard and related globalization
standards. The membership of the consortium represents a broad spectrum
of corporations and organizations in the computer and information
processing industry. Members are: Adobe Systems, Apple, DENIC eG,
Google, Government of India, Government of Tamil Nadu, IBM, Microsoft,
Monotype Imaging, Oracle, SAP, The Society for Natural Language
Technology Research, Sun Microsystems, Sybase, The University of
California at Berkeley, Yahoo!, plus well over a hundred Associate,
Liaison, and Individual members.

For more information, please contact the Unicode Consortium
(http://unicode.org/ <http://www.unicode.org/>).


----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements

Tuesday, June 23, 2009

[Unicode Announcement] Unicode 5.2 beta, updates of UAX #9 and UAX #31

As part of the Unicode 5.2 beta, the following proposed updates for
Unicode Standard Annexes have significant new revisions:

===

UAX #31 Unicode Identifier and Pattern Syntax
http://www.unicode.org/reports/tr31/tr31-10.html

The main changes are the addition of characters and new scripts to
tables for:

* Candidates for Inclusion in Identifiers
* Candidate Characters for Exclusions from Identifiers
* Recommended Scripts

===

UAX#9 Unicode Bidirectional Algorithm
http://www.unicode.org/reports/tr9/tr9-20.html

The main changes are the addition of:

* a section on Bidi Conformance Testing
* the Bidi_Class BN to Rule X6 (removing certain characters from
Bidi processing)
* a clause in HL6 providing for mirroring of R and AL characters in
certain circumstances

===

More details are in the Modifications section of each document, and
there remain some editorial notes asking for feedback on particular
issues. Feedback for both of these documents is solicited by August 3, 2009.

If you have comments for official UTC consideration, please post them by
submitting your comments through our feedback & reporting page:

http://www.unicode.org/reporting.html

If you wish to discuss issues on the Unicode mail list, then please use
the following link to subscribe (if necessary). Please be aware that
discussion comments on the Unicode mail list are not automatically
recorded as input to the UTC. You must use the reporting link above to
generate comments for UTC consideration.

http://www.unicode.org/consortium/distlist.html

----
All of the Unicode Consortium lists are strictly opt-in lists for members
or interested users of our standards. We make every effort to remove
users who do not wish to receive e-mail from us. To see why you are getting
this mail and how to remove yourself from our lists if you want, please
see http://www.unicode.org/consortium/distlist.html#announcements