Language Asset Management (Data Extraction for Alignment)

目次


    What is a language asset?

    There are two types of language assets in this system: “TM (translation memory)” and “Glossary”. In both cases, as a general rule, the source and translation pair (TU = Translation Unit) is counted as one line. You can register (import) your language assets in the following three ways. For details, see User Help on Language Asset List screen. The registered language assets can be used in Quick PE and LAC (Custom MT Model Management) of this system.
    (1) Import the language data (files) that you have locally.
    (2) Extract language data from available texts on the Internet (or files you have locally) by leveraging AI, and import the aligned data.
    (3) Import the stockdata provided in this system.

     

    Service How to use TM
    (Non Stockdata)
    TM
    (Stockdata)
    Glossary Reference  file
    Quick PE Select as a glossary     X  
    Select as a translation memory X   X  
    Use as a reference file X X X X
    Quick MT  Use as a reference file X X X X
    LAC
    (Custom MT Model Management)
    Create a custom trained model by using as corpus for the custom training X X X  
    Create a custom glossary model by setting it to a general/custom MT model     X  
    The language pairs of language assets supported by this system are English-Japanese/Japanese-English only.

    Information in Data Extraction for Alignment screen

    On this screen, you can request to perform the process of extracting language data from texts on the Internet (or files you have locally) using AI. Please set the required fields and click [Request an extraction] button.

    • Extraction From: Select to extract from "Web site" or "File".
    • Name: The name to be given to the extracted data. You can also modify it later.
    • URL: If "Web site" is selected in [Extraction From] dropdown list, enter the target URL. (Example: https://www.k-intl.co.jp/)
    • Select file: If "File" is selected in [Extraction From] dropdown list, browse and select a file you have locally.
    • Language: The language to set for the data for extraction. Set to "en (English)" or "ja (Japanese)".
    • Includes all pages in the Web site: If "Web site" is selected in [Extraction From] dropdown list, when the check box is checked, all pages in the specified site (including the link destinations in the site) are included in the extraction process. If unchecked, only a single page specified by the URL is considered for the extraction process.