Tuesday, January 25, 2011

Save the Date: 35th Internationalization and Unicode Conference: Santa Clara, CA, USA - 10/17-19, 2011

Mountain View, CA, USA - January 25, 2011 - The Unicode® Consortium today announced that the Thirty-fifth Internationalization and Unicode Conference (IUC) will take place in Santa Clara, Calif., USA at the Hyatt Regency Hotel on October 17-19, 2011, sponsored by Adobe. This is the premier conference on technologies and practices for the creation and management of global and multilingual software applications.

This annual event focuses on software and Web globalization. It brings together internationalization experts, tools vendors, software implementers, and business and program managers from around the world. Expert practitioners and industry leaders present detailed recommendations for businesses looking to expand to new international markets and those seeking to improve time to market and cost-efficiency of supporting existing markets. Recent conferences have provided specific advice on designing software for European countries, Latin America, China, India, Japan, Korea, the Middle East, and emerging markets.

This highly rated conference features excellent technical content, industry-tested recommendations and updates on the latest standards and technology. Subject areas include cloud, upgrading to HTML5, integrating with social networking software, and implementing mobile apps. This year's conference will also highlight new features in Unicode Version 6 and other relevant standards published this year.

The Call for Participation will be issued shortly. The abstract submission deadline will be March 25, 2011. For more information about IUC 35, please visit http://www.unicodeconference.org/save-the-date.


About the Event Producer

OMG® is the Event Producer for the Internationalization & Unicode Conferences. OMG is an open membership, not-for-profit consortium that produces and maintains computer industry specifications for interoperable enterprise applications. Our specifications include MDA®, UML®, CORBA®, MOF(tm), XMI® and CWM(tm). OMG's specifications are all available for download by everyone without charge.

For more information about OMG, visit us online at http://www.omg.org.

Note to editors: Unicode Standard, Unicode and the Unicode Logo are trademarks of Unicode, Inc. Unicode Consortium is a registered trademark of Unicode, Inc. OMG and Object Management Group are trademarks of Object Management Group. All other trademarks are the property of their respective owners.

Tuesday, January 18, 2011

Corrected mapping data for Unihan_Variants.txt now available

The Unicode Consortium announces the availability of corrected mapping data for Unihan_Variants.txt. Due to a production problem, many of the traditional and simplified CJK mappings posted in Unihan_Variants.txt as part of the Unihan Database (Unihan.zip) for Unicode 6.0 have corrupted values in them. A corrected version of this mapping data has been posted at: http://www.unicode.org/Public/6.1.0/ucd/Unihan_Variants-6.1.0d1.txt

It is anticipated these corrected mappings will be incorporated in the next version of the Unicode Standard. In the interim, applications which make use of the traditional and simplified CJK mappings in the Unihan Database may wish to correct their mappings based on the revised data file. The traditional and simplified CJK mappings are classified as provisional properties in the Unicode Character Database. Users of provisional properties are cautioned that their use is at the implementer's risk.

Wednesday, January 12, 2011

Unicode Discussion Forum

The Unicode Forum provides a new, more open means for the community of Unicode users and experts to ask questions and discuss topics. For Unicode users, the forum provides an indexed, categorized and easily searchable means of accessing information about the Unicode Standards, related specifications and their use. For experts, the forum provides a place to discuss the desirability and ramification of proposed future extensions to the standard, as well as a place to discuss the state of the art in implementing various features.

To view the forum and participate in discussions, please direct your browser to this URL: http://www.unicode.org/forum/ . Registration is free. As a registered user, you can set up RSS feeds from the forum or subscribe to e-mail notification for topics of interest.

Wednesday, December 29, 2010

Proposed Update UTS #46, Unicode IDNA Compatibility Processing, Version 6.0.1

The Unicode Consortium has released a proposed update for UTS #46, Unicode IDNA Compatibility Processing, Version 6.0.1. This update is intended to make it easier for implementations to both support IDNA2008, and use the mappings in UTS #46. Those mappings allow implementations to meet user expectations for handling uppercase and lowercase, and other character variants, and maintain compatibility with IDNA2003. The proposed data is found in: http://www.unicode.org/Public/idna/6.0.1/

The proposed draft does not change the UTS #46 status or mapping data for Unicode 6.0 characters; instead, it adds new informative fields to the data file and the conformance test file, fields that provide information as to which characters are allowed under IDNA2008. Because UTS #46 is targeted at client software such as browsers, the conformance tests do not check for the CONTEXTO conditions of IDNA2008, which are optional for client software.

Feedback on the proposed draft is welcome. Of particular interest are independent mechanical verification of the new field values, and feedback as to whether it would be useful to add checks for the CONTEXTO conditions to the conformance tests.

Details of the Public Review Issue are on the following web page:

http://www.unicode.org/review/

Review periods for the new items close on January 31, 2011.

If you have comments for official UTC consideration, please post them by submitting your comments through our feedback & reporting page:

http://www.unicode.org/reporting.html

If you wish to discuss issues on the Unicode mail list, then please use the following link to subscribe (if necessary). Please be aware that discussion comments on the Unicode mail list are not automatically recorded as input to the UTC. You must use the reporting link above to generate comments for UTC consideration.

http://www.unicode.org/consortium/distlist.html

Monday, December 20, 2010

Galley proofs for chapters 1-7 of the Unicode 6.0 Core Specification now online

Pre-publication versions of Chapters 1-7 of the Unicode Core Specification, Version 6.0, are now available for online viewing at http://www.unicode.org/versions/Unicode6.0.0/ . These pre-publication chapters are in the final copy editing stage and may have minor edits before the final version is published. The final version of the entire core specification will be published in February 2011.

Thursday, December 2, 2010

Unicode Releases Common Locale Data Repository, Version 1.9

Mountain View, CA, December 1, 2010 - The Unicode® Consortium announced today the release of a new version of the Unicode Common Locale Data Repository (Unicode CLDR 1.9), providing key building blocks for software to support the world's languages. The main features of CLDR 1.9 are enhanced collation and transliteration support, new structure, and modifications for data consistency. The details are found in the CLDR 1.9 Release Note (http://cldr.unicode.org/index/downloads/cldr-1-9).


Unicode CLDR is by far the largest and most extensive standard repository of locale data. This data is used by a wide spectrum of companies for their software internationalization and localization: adapting software to the conventions of different languages for such common software tasks as formatting of dates, times, time zones, numbers, and currency values; sorting text; choosing languages or countries by name; transliterating different alphabets; and many others. Unicode CLDR 1.9 is part of the Unicode locale data project, together with the Unicode Locale Data Markup Language (LDML: http://unicode.org/reports/tr35/). LDML is an XML format used for general interchange of locale data, such as in Microsoft's .NET.


For web pages with different views of CLDR data, see http://unicode.org/cldr/charts.html. For more information about the Unicode CLDR project (including charts) see http://cldr.unicode.org .

Saturday, November 20, 2010

New version of Unicode Ideographic Variation Database released

The Unicode Consortium is pleased to announce the release of version 2010-11-14 of the Unicode Ideographic Variation Database. This release adds a new collection, Hanyo-Denshi, with 4,195 sequences in that collection. Details can be found at <http://www.unicode.org/ivd/> and <http://www.itscj.ipsj.or.jp/domestic/sc02/hanyo-denshi/20100331>.

Friday, November 19, 2010

Corrigendum #8 issued for U+070F SYRIAC ABBREVIATION MARK

Corrigendum #8 has been issued, to correct the Bidi_Class for U+070F SYRIAC ABBREVIATION MARK. This corrigendum corrects the Bidi_Class to the value which was intended for Unicode 6.0, so that U+070F will not be separated into distinct directional runs from the other Syriac characters it is used with.

As for other corrigenda, this correction to the Bidi_Class value for U+070F does not modify the content of Unicode 6.0. However, it makes it possible for applications to declare conformance to Unicode 6.0 plus Corrigendum #8, if needed.

For details please see: http://www.unicode.org/versions/corrigendum8.html http://www.unicode.org/standard/versions/components-6.0.0.html#Unicode_6_0_0_With_Corrigendum