Norsk i den digitale tidsalderen - Meta-Net

More documents

Recommendations

Info

Speech Output Speech Synthesis Speech Input Signal Processing user interface. For static utterances where the wording does not depend on particular contexts of use or per- sonal user data, this can deliver a rich user experience. But more dynamic content in an utterance may suffer from unnatural intonation because different parts of au- dio files have simply been strung together. rough opti- misation, today’s TTS systems are getting better at pro- ducing natural-sounding dynamic utterances. Interfaces in speech interaction have been considerably standardised during the last decade in terms of their various technological components. ere has also been strong market consolidation in speech recognition and speech synthesis. e national markets in the G20 countries (economically resilient countries with high popu- lations) have been dominated by just five global players, with Nuance (USA) and Loquendo (Italy) being the most prominent players in Europe. In 2011, Nuance an- nounced the acquisition of Loquendo, which represents a further step in market consolidation. In the Norwegian TTS market, thirteen Norwegian voices of varying quality are available, some of which have been developed by the above mentioned European players. ree voices have been developed by the Nor- wegian company Lingit, which targets users groups with reading and writing impairments. Another voice was developed by the Norwegian Library of Talking Books and Braille in cooperation with their sister library in Sweden. ere is also an active research community at the Nor- Phonetic Lookup & Intonation Planning Recognition 5: Speech-based dialogue system Natural Language Understanding & Dialogue wegian University of Science and Technology in Trond- heim. Language resources for speech synthesis are abundant for English but only to a lesser extent for Norwegian. e quality of speech synthesis depends heavily on available resources (in particular text corpora tagged for part of speech, tokenisers and pronunciation lexicons) and language specific research on for instance prosodic fea- tures for the language in question. Such resources are abundant for English but only to a lesser extent for Nor- wegian, even if Norwegian is especially challenging due to the wide variety of possible spelling variants and the range of dialects; moreover the tonal accents in most Norwegian dialects and the lack of a one-to-one relation between sounds and letters pose challenges. Regarding dialogue management technology and know-how, the market is rather dominated by national, smaller enterprises. MediaLT has developed a general speech recognition engine used for dialogue management for the visually impaired. Regarding speech-to- text, Max Manus has integrated and localised Philips’ SpeechMagic for Norwegian hospitals. is system is relatively successful, but it is limited to a relatively closed domain (with a closed vocabulary). Recently Dragon Dictation, a voice recognition application for mobile 58
telephones, was launched for Norwegian. is application is the first general dictation system for Norwegian; however, the Norwegian version of Dragon Dictation seems to misinterpret conspicuously more than the En- glish counterpart. Within the domain of speech interaction, a genuine market for the linguistic core technolo- gies for syntactic and semantic analysis does not exist yet. Looking ahead, there will be significant changes, due to the spread of smartphones as a new platform for man- aging customer relationships, in addition to fixed telephones, the Internet and e-mail. is will also affect how speech interaction technology is used. In the long term, there will be fewer telephone-based VUIs, and spoken language apps will play a far more central role as a user-friendly input for smartphones. is will be largely driven by stepwise improvements in the accu- racy of speaker-independent speech recognition via the speech dictation services already offered as centralised services to smartphone users. 4.2.4 Machine Translation e idea of using digital computers to translate natural languages can be traced back to 1946 and was followed by substantial funding for research during the 1950s and again in the 1980s. Yet machine translation (MT) still cannot deliver on its initial promise of providing across- the-board automated translation. e most basic approach to machine translation is the automatic replacement of the words in a text written in one natural language with the equivalent words of another language. is can be useful in subject do- mains that have a very restricted, formulaic language such as weather reports. However, in order to produce a good translation of less restricted texts, larger text units (phrases, sentences, or even whole passages) need to be matched to their closest counterparts in the target language. The major difficulty for Machine Translation is that human language is ambiguous. e major difficulty is that human language is ambiguous. Ambiguity creates challenges on multiple levels, such as word sense disambiguation at the lexical level or syntactic ambiguities at the sentence level. A likely idiomatic translation of the Norwegian sentence Plut- selig røk slangen in English could be Suddenly the hose snapped. However, a simple word-by-word translation of this sentence might yield Suddenly smoked the snake. is is because the verb form røk (past tense of ryke) is ambiguous between ‘snap’ and ‘smoke’ whereas slange is ambiguous between ‘hose’ and ‘snake’; note also that a simple word-by-word translation would not get the dif- ference in word order between Norwegian and English right. In addition to lexical ambiguities and word order dif- ferences, another challenge is syntactic ambiguities. In Norwegian we may topicalise objects, but this possibil- ity is much more restricted in English. e Norwegian sentence Eplene spiste mannen has two possible inter- pretations: either Eplene (the apples) is analysed as the subject of the sentence (the apples ate the man) or as a topicalised object (the apples were eaten by the man). Since this ambiguity does not exist in English, a machine translation system must first identify the correct syntactic interpretation in order to find the correct translation. Another MT challenge for Norwegian is compounding. An efficient translation system must thus be able to dis- cover newly coined compounds, resolve them, and, if needed, create new compounds in the target language. For translations between closely related languages, a translation using direct substitution may be feasible for many sentences. However, rule-based (or linguistic knowledge-driven) systems oen analyse the input text and create an intermediary symbolic representation from which the target language text can be generated. 59
Page 1:
White Paper Series THE NORWEGIAN LA
Page 5 and 6:
FORORD PREFACE Dette dokumentet er
Page 7 and 8:
INNHOLD CONTENTS NORSK I DEN DIGITA
Page 9 and 10:
SAMMENDRAG Informasjonsteknologi p
Page 11 and 12:
er etter hvert mestrer menneskelig
Page 13 and 14:
ke for innbyggerne i et flerspråkl
Page 15 and 16: teknologier som kan overvåke innle
Page 17 and 18: 3 NORSK I DET EUROPEISKE INFORMASJO
Page 19 and 20: grammer er vanligvis ikke dubbet ti
Page 21 and 22: har fått mindre opplæring og vær
Page 23 and 24: Når det gjelder parallell eller ov
Page 25 and 26: Multimedia og multimodale teknologi
Page 27 and 28: Statistisk språkmodell Tekstinput
Page 29 and 30: tekst (eller fonetiske representasj
Page 31 and 32: ligvis også påvirke bruken av tal
Page 33 and 34: ponent som genererer oversatt tekst
Page 35 and 36: kan ekstraheres fra et dokument med
Page 37 and 38: ke ressurser og verktøy til intern
Page 39 and 40: 1. Fremragende støtte 2. God støt
Page 41 and 42: Fremragende God Middels god Fragmen
Page 43: OM META-NET META-NET er et forsknin
Page 46 and 47: technology will feature soware that
Page 48 and 49: LANGUAGES AT RISK: A CHALLENGE FOR
Page 50 and 51: tion. According to one estimate, th
Page 52 and 53: Humans acquire language skills in t
Page 54 and 55: Norwegian’s wide variety of diale
Page 56 and 57: way has more wide-ranging responsib
Page 58 and 59: ates take place on the Internet Dig
Page 60 and 61: LANGUAGE TECHNOLOGY SUPPORT FOR NOR
Page 62 and 63: Input Text Output Pre-processing Gr
Page 64 and 65: Web Pages Pre-processing Semantic P
Page 68 and 69: Statistical Machine Translation Sou
Page 70 and 71: languages that benefit from a consi
Page 72 and 73: knowledge and expertise in academia
Page 74 and 75: uantity Availability uality Languag
Page 76 and 77: them openly available to the resear
Page 78 and 79: Excellent Good Moderate Fragmentary
Page 81 and 82: A LITTERATURLISTE REFERENCES [1] Al
Page 83: [23] Jerrold H. Zar. Candidate for
Page 86 and 87: Malta Malta Department Intelligent
Page 88 and 89: Omtrent 100 eksperter innen språkt
Page 90: In everyday communication, Europe
show all

Norsk i den digitale tidsalderen - Meta-Net

You also want an ePaper? Increase the reach of your titles

Delete template?

Save as template?