MESA And SMPTE Develop First Human- And Machine-Readable Language Metadata Table
In collaboration with MESA, SMPTE is bringing the media industry’s first human- and machine-readable Language Metadata Table (LMT) into its public review process, an early step toward SMPTE standardization
Published as a SMPTE Public Committee Draft (CD), the vetted and approved list of language codes will be readily available for public comment, implementation, and validation.
The LMT register is intended to give media companies, content owners, video service providers, and others a controlled vocabulary and standardized set of codes for accurately and consistently identifying spoken and written language, in turn supporting more efficient interchange of media worldwide. Reflecting many thousands of permutations, LMT codes support numerous applications including audio, written and timed text (closed captions and subtitles), accessibility, licensing, content localization, and international distribution.
“The Language Metadata Table was started at WarnerMedia in 2017 to normalize language codes within the organization, and IETF BCP 47 was selected due to its flexibility,” said Yonah Levenson, LMT chair at MESA, the LMT sponsor. “As interest in LMT as the M&E industry's language code solution increased, SMPTE recognized the value of the LMT and came on board as the technical partner/advisor.”
Work on the LMT register has begun in SMPTE Technology Committees (TCs), which will produce a SMPTE Public CD in the first half of 2021. The Public CD process allows SMPTE to put the LMT register into the public domain quickly and then start the work of gathering feedback and making improvements to both the register and guidelines for its independent management by multiple stakeholders.
“If you buy and sell media, you understand that a common vocabulary for language tagging is sorely needed,” said SMPTE Standards VP Bruce Devlin. “The LMT register accounts for all languages as well as dialects and scripts. As we see the register through the Public CD process, our hope is that the LMT register will become a canonical resource that serves the needs of all media organizations and ecosystems. Accessing this data will be as simple as clicking on a link or using an API to grab required codes.”
SMPTE TCs are reviewing the prototype LMT register to determine if the structure of the dictionary is correct and if the process for updating that dictionary is correct. After this step is complete, the dictionary and update process will enter a public review period, during which people and organizations can try out the register and use a dedicated GitHub repository at github.com/smpte to provide real-world feedback that will inform iterative improvement of the register. The Society will leverage the SMPTE Knowledge Network, which is built on a flexible Microsoft Teams environment with integrated apps including the Microsoft 365 suite and GitHub, to bring agility and efficiency to the Public CD process.
“My hope is that ultimately we will have a structure document in SMPTE that defines the LMT, presents the data itself in both human- and machine-readable form, and provides a new administrative guideline that describes how we’ll manage controlled vocabularies and ontologies for third parties. I encourage any organization or individual with a stake in the internationalization of content to join the appropriate SMPTE TC and contribute their requirements and expertise,” added Devlin.
You might also like...
HDR & WCG For Broadcast: Part 3 - Achieving Simultaneous HDR-SDR Workflows
Welcome to Part 3 of ‘HDR & WCG For Broadcast’ - a major 10 article exploration of the science and practical applications of all aspects of High Dynamic Range and Wide Color Gamut for broadcast production. Part 3 discusses the creative challenges of HDR…
IP Security For Broadcasters: Part 4 - MACsec Explained
IPsec and VPN provide much improved security over untrusted networks such as the internet. However, security may need to improve within a local area network, and to achieve this we have MACsec in our arsenal of security solutions.
Standards: Part 23 - Media Types Vs MIME Types
Media Types describe the container and content format when delivering media over a network. Historically they were described as MIME Types.
Building Software Defined Infrastructure: Part 1 - System Topologies
Welcome to Part 1 of Building Software Defined Infrastructure - a new multi-part content collection from Tony Orme. This series is for broadcast engineering & IT teams seeking to deepen their technical understanding of the microservices based IT technologies that are…
IP Security For Broadcasters: Part 3 - IPsec Explained
One of the great advantages of the internet is that it relies on open standards that promote routing of IP packets between multiple networks. But this provides many challenges when considering security. The good news is that we have solutions…