English
Related papers

Related papers: Japanese word sense disambiguation based on exampl…

200 papers

It is challenging to translate names and technical terms across languages with different alphabets and sound inventories. These items are commonly transliterated, i.e., replaced with approximate phonetic equivalents. For example, "computer"…

cmp-lg · Computer Science 2009-10-02 Kevin Knight , Jonathan Graehl

The most commonly used Japanese alphabets are Kanji, Hiragana and Katakana. The Kanji alphabet includes pictographs or ideographic characters that were adopted from the Chinese alphabet. Hiragana is used to spell words of Japanese origin,…

Human-Computer Interaction · Computer Science 2013-10-14 Umakant Mishra

Sign language is a visual language expressed through hand movements and non-manual markers. Non-manual markers include facial expressions and head movements. These expressions vary across different nations. Therefore, specialized analysis…

Computer Vision and Pattern Recognition · Computer Science 2025-06-25 Yui Tatsumi , Shoko Tanaka , Shunsuke Akamatsu , Takahiro Shindo , Hiroshi Watanabe

An evaluation of distributed word representation is generally conducted using a word similarity task and/or a word analogy task. There are many datasets readily available for these tasks in English. However, evaluating distributed…

Computation and Language · Computer Science 2018-02-23 Yuya Sakaizawa , Mamoru Komachi

We provide a detailed overview of various approaches to word segmentation of Asian Languages, specifically Chinese, Korean, and Japanese languages. For each language, approaches to deal with word segmentation differs. We also include our…

Computation and Language · Computer Science 2024-07-30 Matthew Rho , Yexin Tian , Qin Chen

Word sense disambiguation has developed as a sub-area of natural language processing, as if, like parsing, it was a well-defined task which was a pre-requisite to a wide range of language-understanding applications. First, I review earlier…

cmp-lg · Computer Science 2007-05-23 Adam Kilgarriff

Japan is a unique country with a distinct cultural heritage, which is reflected in billions of historical documents that have been preserved. However, the change in Japanese writing system in 1900 made these documents inaccessible for the…

Computation and Language · Computer Science 2021-06-15 Alex Lamb , Tarin Clanuwat , Siyu Han , Mikel Bober-Irizar , Asanobu Kitamoto

Particles fullfill several distinct central roles in the Japanese language. They can mark arguments as well as adjuncts, can be functional or have semantic funtions. There is, however, no straightforward matching from particles to…

Computation and Language · Computer Science 2007-05-23 Melanie Siegel

Word sense disambiguation assumes word senses. Within the lexicography and linguistics literature, they are known to be very slippery entities. The paper looks at problems with existing accounts of `word sense' and describes the various…

cmp-lg · Computer Science 2007-05-23 Adam Kilgarriff

Possessive pronouns are used as determiners in English when no equivalent would be used in a Japanese sentence with the same meaning. This paper proposes a heuristic method of generating such possessive pronouns even when there is no…

cmp-lg · Computer Science 2008-02-03 Francis Bond , Kentaro Ogura , Satoru Ikehara

Recent years have seen an increase in the number of large-scale multilingual NLP projects. However, even in such projects, languages with special processing requirements are often excluded. One such language is Japanese. Japanese is written…

Computation and Language · Computer Science 2020-10-15 Paul McCann

If someone is looking for a certain publication in the field of computer science, the searching person is likely to use the DBLP to find the desired publication. The DBLP data set is continuously extended with new publications, or rather…

Computation and Language · Computer Science 2017-09-27 Paul Christian Sommerhoff

The important part of semantics of complex sentence is captured as relations among semantic roles in subordinate and main clause respectively. However if there can be relations between every pair of semantic roles, the amount of computation…

cmp-lg · Computer Science 2008-02-03 Hiroshi Nakagawa , Shin'ichiro Nishizawa

We present a broad coverage Japanese grammar written in the HPSG formalism with MRS semantics. The grammar is created for use in real world applications, such that robustness and performance issues play an important role. It is connected to…

Computation and Language · Computer Science 2007-05-23 Melanie Siegel , Emily M. Bender

This study proposes a method to develop neural models of the morphological analyzer for Japanese Hiragana sentences using the Bi-LSTM CRF model. Morphological analysis is a technique that divides text data into words and assigns information…

Computation and Language · Computer Science 2022-01-11 Jun Izutsu , Kanako Komiya

We report the development of Japanese SimCSE, Japanese sentence embedding models fine-tuned with SimCSE. Since there is a lack of sentence embedding models for Japanese that can be used as a baseline in sentence embedding research, we…

Computation and Language · Computer Science 2023-10-31 Hayato Tsukagoshi , Ryohei Sasano , Koichi Takeda

Lexical complexity prediction (LCP) is the task of predicting the complexity of words in a text on a continuous scale. It plays a vital role in simplifying or annotating complex words to assist readers. To study lexical complexity in…

Computation and Language · Computer Science 2023-07-03 Yusuke Ide , Masato Mita , Adam Nohejl , Hiroki Ouchi , Taro Watanabe

Text-based communication, such as text chat, is commonly employed in various contexts, both professional and personal. However, it lacks the rich emotional cues present in verbal and visual forms of communication, such as facial expressions…

Human-Computer Interaction · Computer Science 2023-09-14 Rintaro Chujo , Atsunobu Suzuki , Ari Hautasaari

We evaluate a battery of recent large language models on two benchmarks for word sense disambiguation in Swedish. At present, all current models are less accurate than the best supervised disambiguators in cases where a training set is…

Computation and Language · Computer Science 2024-10-31 Richard Johansson

In Japanese text-to-speech (TTS), it is necessary to add accent information to the input sentence. However, there are a limited number of publicly available accent dictionaries, and those dictionaries e.g. UniDic, do not contain many…

Computation and Language · Computer Science 2020-09-22 Hideyuki Tachibana , Yotaro Katayama
‹ Prev 1 2 3 10 Next ›