EnglishFrenchSpanish

Ad


OnWorks favicon

make-lingua-identify-languagep - Online in the Cloud

Run make-lingua-identify-languagep in OnWorks free hosting provider over Ubuntu Online, Fedora Online, Windows online emulator or MAC OS online emulator

This is the command make-lingua-identify-languagep that can be run in the OnWorks free hosting provider using one of our multiple free online workstations such as Ubuntu Online, Fedora Online, Windows online emulator or MAC OS online emulator

PROGRAM:

NAME


make-lingua-identify-language - creates language modules for Lingua::Identify

SYNOPSIS


make-lingua-identify-language Language-Tag Language-Name file1 [file2 ...]

or

make-lingua-identify-language -d TAG1-LANGUAGE1/ [TAG2-LANGUAGE2/ ...]

or

make-lingua-identify

DESCRIPTION


Creates language modules to be used by Lingua::Identify.

After creating the modules, you still have to install them.

Please note that this script is still at an early stage. Please do not even look at the
code...

Without parameters, make-lingua-identify-language assumes mode -d and goes through all the
directories in the current one. This is useful to be used in a directory where you
something like this:

.
|-- en-english
| `-- english.txt
|-- fr-french
| `-- french1.txt
| `-- french2.txt
`-- pt-portuguese
`-- portuguese.txt

OPTIONS
-d
Directory mode. Each parameter passed should be a directory whose name must be of the form
tag-name (e.g., en-english/ ). Each of the directories passed should contain text files
that can be used to train Lingua::Identify.

-D
Debug mode. Only for development.

-h
Display help and exit.

-v
Show version and exit.

-verbose
Verbose mode.

-locale="<locale>"
Set a specific locale. This way your text will be all lowercased before analysed.

META.yml


"META.yml" files are not parsed as other files, they are ignored.

In directory mode ("-d" switch), "META.yml" files are checked for info on languages codes
and sets.

Here's a simple "META.yml" for you to put in your directories:

two_letter_code: pt
three_letter_code: por
sets:
spoken_in_portugal

With that, the language will be identified with the two letter code "pt" or the three
letter code "por"; it will also be in the set ":spoken_in_portugal".

CONTRIBUTING WITH NEW LANGUAGES


Please do not contribute with modules you made yourself. It's easier to contribute with
unprocessed text, because that allows for new versions of Lingua::Identify not having to
drop languages down in case I can't contact you by that time.

Use make-lingua-identify-language to create a new module for your own personal use, if you
must, but try to contribute with unprocessed text rather than those modules.

Use make-lingua-identify-languagep online using onworks.net services


Free Servers & Workstations

Download Windows & Linux apps

  • 1
    Phaser
    Phaser
    Phaser is a fast, free, and fun open
    source HTML5 game framework that offers
    WebGL and Canvas rendering across
    desktop and mobile web browsers. Games
    can be co...
    Download Phaser
  • 2
    VASSAL Engine
    VASSAL Engine
    VASSAL is a game engine for creating
    electronic versions of traditional board
    and card games. It provides support for
    game piece rendering and interaction,
    and...
    Download VASSAL Engine
  • 3
    OpenPDF - Fork of iText
    OpenPDF - Fork of iText
    OpenPDF is a Java library for creating
    and editing PDF files with a LGPL and
    MPL open source license. OpenPDF is the
    LGPL/MPL open source successor of iText,
    a...
    Download OpenPDF - Fork of iText
  • 4
    SAGA GIS
    SAGA GIS
    SAGA - System for Automated
    Geoscientific Analyses - is a Geographic
    Information System (GIS) software with
    immense capabilities for geodata
    processing and ana...
    Download SAGA GIS
  • 5
    Toolbox for Java/JTOpen
    Toolbox for Java/JTOpen
    The IBM Toolbox for Java / JTOpen is a
    library of Java classes supporting the
    client/server and internet programming
    models to a system running OS/400,
    i5/OS, o...
    Download Toolbox for Java/JTOpen
  • 6
    D3.js
    D3.js
    D3.js (or D3 for Data-Driven Documents)
    is a JavaScript library that allows you
    to produce dynamic, interactive data
    visualizations in web browsers. With D3
    you...
    Download D3.js
  • More »

Linux commands

  • 1
    abidiff
    abidiff
    abidiff - compare ABIs of ELF files
    abidiff compares the Application Binary
    Interfaces (ABI) of two shared libraries
    in ELF format. It emits a meaningful
    repor...
    Run abidiff
  • 2
    abidw
    abidw
    abidw - serialize the ABI of an ELF
    file abidw reads a shared library in ELF
    format and emits an XML representation
    of its ABI to standard output. The
    emitted ...
    Run abidw
  • 3
    copac2xml
    copac2xml
    bibutils - bibliography conversion
    utilities ...
    Run copac2xml
  • 4
    copt
    copt
    copt - peephole optimizer SYSNOPIS:
    copt file.. DESCRIPTION: copt is a
    general-purpose peephole optimizer. It
    reads code from its standard input and
    writes an ...
    Run copt
  • 5
    gather_stx_titles
    gather_stx_titles
    gather_stx_titles - gather title
    declarations from Stx documents ...
    Run gather_stx_titles
  • 6
    gatling-bench
    gatling-bench
    bench - http benchmark ...
    Run gatling-bench
  • More »

Ad