Compact and Robust Models for Japanese-English Character-level Machine Translation

Jinan Dai; Kazunori Yamaguchi

doi:10.18653/v1/D19-5202

Compact and Robust Models for Japanese-English Character-level Machine Translation

Correct Metadata for

Use this form to create a GitHub issue with structured data describing the correction. You will need a GitHub account. Once you create that issue, the correction will be reviewed by a staff member.

⚠️ Mobile Users: Submitting this form to create a new issue will only work with github.com, not the GitHub Mobile app.

Important: The Anthology treat PDFs as authoritative. Please use this form only to correct data that is out of line with the PDF. See our corrections guidelines if you need to change the PDF.

Title Adjust the title. Retain tags such as <fixed-case>.

Authors Adjust author names and order to match the PDF.

Abstract Correct abstract if needed. Retain XML formatting tags such as <tex-math>. You may use <b>...</b> for bold, <i>...</i> for italic, and <url>...</url> for URLs.

Verification against PDF Ensure that the new title/authors match the snapshot below. (If there is no snapshot or it is too small, consult the PDF.)

Authors concatenated from the text boxes above:

ALL author names match the snapshot above—including middle initials, hyphens, and accents.

Abstract

Character-level translation has been proved to be able to achieve preferable translation quality without explicit segmentation, but training a character-level model needs a lot of hardware resources. In this paper, we introduced two character-level translation models which are mid-gated model and multi-attention model for Japanese-English translation. We showed that the mid-gated model achieved the better performance with respect to BLEU scores. We also showed that a relatively narrow beam of width 4 or 5 was sufficient for the mid-gated model. As for unknown words, we showed that the mid-gated model could somehow translate the one containing Katakana by coining out a close word. We also showed that the model managed to produce tolerable results for heavily noised sentences, even though the model was trained with the dataset without noise.

Anthology ID:: D19-5202
Volume:: Proceedings of the 6th Workshop on Asian Translation
Month:: November
Year:: 2019
Address:: Hong Kong, China
Editors:: Toshiaki Nakazawa, Chenchen Ding, Raj Dabre, Anoop Kunchukuttan, Nobushige Doi, Yusuke Oda, Ondřej Bojar, Shantipriya Parida, Isao Goto, Hidaya Mino
Venue:: WAT
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 36–44
Language:
URL:: https://aclanthology.org/D19-5202/
DOI:: 10.18653/v1/D19-5202
Bibkey:
Cite (ACL):: Jinan Dai and Kazunori Yamaguchi. 2019. Compact and Robust Models for Japanese-English Character-level Machine Translation. In Proceedings of the 6th Workshop on Asian Translation, pages 36–44, Hong Kong, China. Association for Computational Linguistics.
Cite (Informal):: Compact and Robust Models for Japanese-English Character-level Machine Translation (Dai & Yamaguchi, WAT 2019)
Copy Citation:
PDF:: https://aclanthology.org/D19-5202.pdf

PDF Cite Search Fix data

Export citation

BibTeX
MODS XML
Endnote
Preformatted

@inproceedings{dai-yamaguchi-2019-compact,
    title = "Compact and Robust Models for {J}apanese-{E}nglish Character-level Machine Translation",
    author = "Dai, Jinan  and
      Yamaguchi, Kazunori",
    editor = "Nakazawa, Toshiaki  and
      Ding, Chenchen  and
      Dabre, Raj  and
      Kunchukuttan, Anoop  and
      Doi, Nobushige  and
      Oda, Yusuke  and
      Bojar, Ond{\v{r}}ej  and
      Parida, Shantipriya  and
      Goto, Isao  and
      Mino, Hidaya",
    booktitle = "Proceedings of the 6th Workshop on Asian Translation",
    month = nov,
    year = "2019",
    address = "Hong Kong, China",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/D19-5202/",
    doi = "10.18653/v1/D19-5202",
    pages = "36--44",
    abstract = "Character-level translation has been proved to be able to achieve preferable translation quality without explicit segmentation, but training a character-level model needs a lot of hardware resources. In this paper, we introduced two character-level translation models which are mid-gated model and multi-attention model for Japanese-English translation. We showed that the mid-gated model achieved the better performance with respect to BLEU scores. We also showed that a relatively narrow beam of width 4 or 5 was sufficient for the mid-gated model. As for unknown words, we showed that the mid-gated model could somehow translate the one containing Katakana by coining out a close word. We also showed that the model managed to produce tolerable results for heavily noised sentences, even though the model was trained with the dataset without noise."
}

Download as File

<?xml version="1.0" encoding="UTF-8"?>
<modsCollection xmlns="http://www.loc.gov/mods/v3">
<mods ID="dai-yamaguchi-2019-compact">
    <titleInfo>
        <title>Compact and Robust Models for Japanese-English Character-level Machine Translation</title>
    </titleInfo>
    <name type="personal">
        <namePart type="given">Jinan</namePart>
        <namePart type="family">Dai</namePart>
        <role>
            <roleTerm authority="marcrelator" type="text">author</roleTerm>
        </role>
    </name>
    <name type="personal">
        <namePart type="given">Kazunori</namePart>
        <namePart type="family">Yamaguchi</namePart>
        <role>
            <roleTerm authority="marcrelator" type="text">author</roleTerm>
        </role>
    </name>
    <originInfo>
        <dateIssued>2019-11</dateIssued>
    </originInfo>
    <typeOfResource>text</typeOfResource>
    <relatedItem type="host">
        <titleInfo>
            <title>Proceedings of the 6th Workshop on Asian Translation</title>
        </titleInfo>
        <name type="personal">
            <namePart type="given">Toshiaki</namePart>
            <namePart type="family">Nakazawa</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Chenchen</namePart>
            <namePart type="family">Ding</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Raj</namePart>
            <namePart type="family">Dabre</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Anoop</namePart>
            <namePart type="family">Kunchukuttan</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Nobushige</namePart>
            <namePart type="family">Doi</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Yusuke</namePart>
            <namePart type="family">Oda</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Ondřej</namePart>
            <namePart type="family">Bojar</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Shantipriya</namePart>
            <namePart type="family">Parida</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Isao</namePart>
            <namePart type="family">Goto</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <name type="personal">
            <namePart type="given">Hidaya</namePart>
            <namePart type="family">Mino</namePart>
            <role>
                <roleTerm authority="marcrelator" type="text">editor</roleTerm>
            </role>
        </name>
        <originInfo>
            <publisher>Association for Computational Linguistics</publisher>
            <place>
                <placeTerm type="text">Hong Kong, China</placeTerm>
            </place>
        </originInfo>
        <genre authority="marcgt">conference publication</genre>
    </relatedItem>
    <abstract>Character-level translation has been proved to be able to achieve preferable translation quality without explicit segmentation, but training a character-level model needs a lot of hardware resources. In this paper, we introduced two character-level translation models which are mid-gated model and multi-attention model for Japanese-English translation. We showed that the mid-gated model achieved the better performance with respect to BLEU scores. We also showed that a relatively narrow beam of width 4 or 5 was sufficient for the mid-gated model. As for unknown words, we showed that the mid-gated model could somehow translate the one containing Katakana by coining out a close word. We also showed that the model managed to produce tolerable results for heavily noised sentences, even though the model was trained with the dataset without noise.</abstract>
    <identifier type="citekey">dai-yamaguchi-2019-compact</identifier>
    <identifier type="doi">10.18653/v1/D19-5202</identifier>
    <location>
        <url>https://aclanthology.org/D19-5202/</url>
    </location>
    <part>
        <date>2019-11</date>
        <extent unit="page">
            <start>36</start>
            <end>44</end>
        </extent>
    </part>
</mods>
</modsCollection>

Download as File

%0 Conference Proceedings
%T Compact and Robust Models for Japanese-English Character-level Machine Translation
%A Dai, Jinan
%A Yamaguchi, Kazunori
%Y Nakazawa, Toshiaki
%Y Ding, Chenchen
%Y Dabre, Raj
%Y Kunchukuttan, Anoop
%Y Doi, Nobushige
%Y Oda, Yusuke
%Y Bojar, Ondřej
%Y Parida, Shantipriya
%Y Goto, Isao
%Y Mino, Hidaya
%S Proceedings of the 6th Workshop on Asian Translation
%D 2019
%8 November
%I Association for Computational Linguistics
%C Hong Kong, China
%F dai-yamaguchi-2019-compact
%X Character-level translation has been proved to be able to achieve preferable translation quality without explicit segmentation, but training a character-level model needs a lot of hardware resources. In this paper, we introduced two character-level translation models which are mid-gated model and multi-attention model for Japanese-English translation. We showed that the mid-gated model achieved the better performance with respect to BLEU scores. We also showed that a relatively narrow beam of width 4 or 5 was sufficient for the mid-gated model. As for unknown words, we showed that the mid-gated model could somehow translate the one containing Katakana by coining out a close word. We also showed that the model managed to produce tolerable results for heavily noised sentences, even though the model was trained with the dataset without noise.
%R 10.18653/v1/D19-5202
%U https://aclanthology.org/D19-5202/
%U https://doi.org/10.18653/v1/D19-5202
%P 36-44

Download as File

Markdown (Informal)

[Compact and Robust Models for Japanese-English Character-level Machine Translation](https://aclanthology.org/D19-5202/) (Dai & Yamaguchi, WAT 2019)

Compact and Robust Models for Japanese-English Character-level Machine Translation (Dai & Yamaguchi, WAT 2019)

ACL

Jinan Dai and Kazunori Yamaguchi. 2019. Compact and Robust Models for Japanese-English Character-level Machine Translation. In Proceedings of the 6th Workshop on Asian Translation, pages 36–44, Hong Kong, China. Association for Computational Linguistics.