Thursday, January 21, 2016

Proposal to Remove Some Hira/Kata From Script_Extensions

The Script_Extensions property values for some characters contain Hiragana, Katakana, or Bopomofo, when they should only contain Han. The Unicode Technical Committee is considering removing the Hiragana, Katakana, or Bopomofo in these cases, and would like feedback as to any that should not be changed, and any others that should be. Public Review Issue #316 contains details of a proposal to remove these items from Script_Extensions.

For information about how to discuss this Public Review Issue and how to supply formal feedback, please see the feedback and discussion instructions.

Thursday, January 14, 2016

Proposed Update UAX #9, Unicode Bidirectional Algorithm

A new proposed update of UAX #9, Unicode Bidirectional Algorithm for the Unicode 9.0 release is now available for public review and comment.

The table in Section 2.7, Markup and Formatting, has been updated to reflect changes to isolates in HTML5 and CSS.

For further information and instructions on how to leave feedback, please see Public Review Issue #315.

Proposed Update UAX #45, U-Source Ideographs

A new proposed update of UAX #45, U-Source Ideographs for the Unicode 9.0 release is now available for public review and comment.

Many updates and additions have been made to the USourceData.txt and the accompanying list of glyphs for all the U-Source ideographs, USourceGlyphs.pdf. For the latest versions of the source data and glyph files for review, see the versioned files posted in the Unicode 9.0 UCD data file review directory.

For further information and instructions on how to leave feedback, please see Public Review Issue #314.

Wednesday, January 13, 2016

Proposed Update UTS #39, Unicode Security Mechanisms

A new proposed update of UTS #39, Unicode Security Mechanisms is now available for public review and comment.

The proposed update of this Unicode Technical Standard includes new material for email security profiles and text about the use of Script_Extensions. The data file confusablesWholeScript.txt has been withdrawn, because in practice the process of derivation of whole script confusables depends on the particular set of characters supported by an application. The use of a data file is replaced by a logical process of deriving the whole-script confusables data based on the set of supported characters.

For further information and instructions on how to leave feedback, please see Public Review Issue #313.

Wednesday, January 6, 2016

Unicode Tutorial Workshop in Oman (Feb 14-16, 2016)


This tutorial workshop, sponsored by the Unicode Consortium and organized by the German University of Technology in Oman, is a three-day event designed to familiarize the audience with the Unicode Standard and the concepts of internationalization. It is the first ever Unicode event to be held in the Middle East.

The workshop program includes an introduction to Writing Systems & Unicode, plus presentations on Arabic Typography, web best practices, mobile internationalization, and more.

The workshop website provides full information about the event.

Tuesday, January 5, 2016

Feedback on Draft additional repertoire for ISO/IEC 10646:2016 (5th edition) CD2


The Unicode Technical Committee is soliciting feedback on pending additions to the draft repertoire of characters, to help discover any errors in character names, incorrect glyphs, or other problems. There is a short window of opportunity to review and comment on the repertoire additions noted below.

The following additional repertoire from ISO/IEC 10646:2016 (5th Edition), which is in committee ballot, is under review. See the associated repertoire in: Draft additional repertoire for ISO/IEC 10646:2016 (5th edition) CD2.

The Unicode Standard is developed in synchrony with ISO/IEC 10646. After ISO balloting is completed on any repertoire additions, no further changes or corrections will be possible. (See the FAQ Standards Developing Organizations for additional information on the stages in ISO standards development.) Advance feedback on these repertoire additions will help inform the UTC discussions about its own contribution to the ISO balloting process.

Documents referenced in the draft repertoire with numbers such as L2/15-088 are available in the UTC Document Registry.

For information about how to discuss this Public Review Issue and how to supply formal feedback, please see the feedback and discussion instructions.

Monday, December 21, 2015

Unicode Updates Emoji Charts and Expands List of Candidates

U+1F381 WRAPPED PRESENT ImageThe Unicode Consortium announced today it is updating the Unicode emoji charts with the following major changes:

The Full Emoji Data chart adds the latest updates for Google, Twitter, Windows, and EmojiOne emoji images. It also now includes all the emoji sequences, for a total of 1,624 emoji.

The Emoji Candidates chart is updated to add the 7 characters recently approved as candidates, for a total of 74 emoji candidates. There are refreshed images courtesy of Adobe, Emojipedia.org and EmojiXpress.

The Emoji Ordering chart now shows the diverse family emoji, and adds subcategories to make the organization clearer.

And the Emoji Style chart has been extended to list the 1,624 emoji characters and sequences as text with variation selectors and fonts, for testing with browsers.

Show your support of Unicode, and adopt a character!

Wednesday, December 16, 2015

Unicode Launches Adopt-a-Character Campaign to Support the World’s “Digitally Disadvantaged” Living Languages

Non-profit consortium invites public to adopt any emoji, letter or symbol as fun, meaningful gifts that fund research and coding needed to support minority languages

U+1F381 WRAPPED PRESENT ImageMOUNTAIN VIEW, Calif.—(BUSINESS WIRE)—Unicode Consortium, the 501(c)(3) non-profit that standardizes the way computers represent text in all languages – including emoji characters – today announced its Adopt-a-Character campaign. The new program is an opportunity to adopt and dedicate an emoji, letter or any symbol on the keyboard to help Unicode’s important work of supporting the world’s languages in digital form. Adoption options are available at $100, $1,000 and $5,000 levels and make meaningful and fun gifts for the holidays or any occasion. Adoption donations are tax deductible in the U.S.

Funds raised will be used to support Unicode’s core mission of developing and extending the necessary standards, data and software to support the world’s living languages. Unicode works with linguists, experts, cultural leaders and technologists to create coding standards to support minority languages in digital form.

“Beyond our work standardizing emoji, Unicode is tackling some big challenges that might surprise many people,” said Mark Davis, co-founder and president of the Unicode Consortium and an internationalization expert at Google. “The vast majority of the world’s living languages, close to 98 percent, are ‘digitally disadvantaged’ – meaning they are not supported on the most popular devices, operating systems, browsers and mobile applications. For example, only a handful of African languages have adequate digital support. The funds from our new Adopt-a-Character campaign will help us continue the important standardization work that is best done by a neutral organization like Unicode.”

Ensuring Digital Vitality, from Cherokee to N’Ko

So far, Unicode’s resources have been focused on the most-prominent scripts and languages of the world. Gathering information for less-prominent scripts and languages – such as Berber, Balinese, Cherokee, Javanese, N’Ko, Pahawh Hmong and Kashmiri – is often more difficult, requiring travel, research, engineering resources and software tooling.

Just 15 years ago, Cherokee was not available digitally and now as a result of Unicode’s work it can be found on computers, mobile devices such as the iPhone and iPad, and on Gmail. Because of Unicode’s work standardizing N’Ko – a script used to write a number of the West African Mande languages, with a population of over 20 million people – publishers are now able to modernize their operations, print in multiple locations and reach a broader audience.

“The Internet has made us all more acutely aware of how small our world is and how rich the creations of its inhabitants are,” said Greg Welch, a Unicode board member and Senior Director, Strategic Marketing, Mobile Client Platforms at Intel. “As we become a more connected and paperless global society, we cannot leave minority and digitally disadvantaged languages behind. It’s vital to ensure that the text on which a culture’s propagation depends makes it across the digital divide.”

How to Adopt-a-Character

More information about Adopt-a-Character can be found at http://unicode.org/consortium/adopt-a-character.html

About Unicode Consortium

The Unicode Consortium’s mission is to lay a solid foundation for digital support of the world’s languages. If you've used any computer or smartphone, then you're using Unicode and have benefited from the consortium’s work. The consortium – whose members include companies such as Adobe, Apple, Facebook, Google, IBM, Microsoft and more – is a 501(c)(3) non-profit that emerged from the technology industry’s effort to standardize the way computers represent text (including emoji) in all languages – from English to Chinese to Zulu – across different devices and operating systems. The group operates largely as a volunteer organization that is funded by membership fees and donations. A full list of members is on http://unicode.org/consortium/members.html