Send Moses-support mailing list submissions to
moses-support@mit.edu
To subscribe or unsubscribe via the World Wide Web, visit
http://mailman.mit.edu/mailman/listinfo/moses-support
or, via email, send a message with subject or body 'help' to
moses-support-request@mit.edu
You can reach the person managing the list at
moses-support-owner@mit.edu
When replying, please edit your Subject line so it is more specific
than "Re: Contents of Moses-support digest..."
Today's Topics:
1. Fwd: TweetMT @ SEPLN 2015 (Cristina)
2. Chinese segmentation/tokenization (Marcin Junczys-Dowmunt)
3. CALL FOR BOOK PROPOSALS (Rohit Gupta)
4. Re: Problem (Hieu Hoang)
----------------------------------------------------------------------
Message: 1
Date: Fri, 20 Mar 2015 11:40:12 +0100
From: Cristina <cristinae@lsi.upc.edu>
Subject: [Moses-support] Fwd: TweetMT @ SEPLN 2015
To: moses-support@mit.edu
Message-ID:
<CAL0MP8iORLVAb08AYbLdrQRX+Kk_+k-pAjDuTy+QO5xGkuv83Q@mail.gmail.com>
Content-Type: text/plain; charset="utf-8"
Apologies for multiple postings
*************************************************************************TweetMT
2015--Tweet Translation Workshop at SEPLN 2015
TweetMT is a workshop and shared task on machine translation applied to
tweets. It will take place in September, 2015, in Alicante, co-located with
SEPLN 2015 (to be confirmed). The objective of the task is to bring
together interested researchers to join forces to experiment with and
compare different approaches to tweet MT. This workshop is a follow-up to
two other workshops organized previously also at SEPLN: TweetNorm2013 and
TweetLID2014.
The machine translation of tweets is a complex task that greatly depends on
the type of data we work with. The translation process of tweets is very
different from that of correct texts posted for instance through a content
manager. Tweets are often written from mobile devices, which exacerbates
the poor quality of the spelling, and include errors, symbols and
diacritics. The texts also vary in terms of structure, where the latter
include tweet-specific features such as hashtags, user mentions, and
retweets, among others. The translation of tweets can be tackled as a
direct translation (tweet-to-tweet) or as an indirect translation (tweet
normalization to standard text (Kaufmann&Kalita, 2011), text translation
and, if needed, tweet generation). Although the first approach looks
attractive, the lack of parallel or comparable tweets for the working
languages (Petrovic et al., 2010) tends to lead us towards an indirect
approach. Some authors also try to gather similar tweets in other languages
(CLIR).
Work in this area is scarce in the literature but a growing interest is
evident (Gotti et al., 2013). An important point of reference is the work
done to translate SMS texts during the Haiti earthquake (Munro, 2010).
The current task will focus on MT of tweets between languages of the
Iberian Peninsula (Basque, Catalan, Galician, Portuguese and Spanish), as
well as English. The organizing committee will release development data
including parallel tweets that will enable participants to train their
systems. For the final evaluation participants will have to submit the
automatic translation of a number of tweet corpora in a short period of
time. The evaluation will be carried out using automatic distances to the
reference corpora.
These corpora are not meant to be representative of all types of messages
that can be observed in informal communication. This is instead an initial
attempt at tackling part of the task which starts by addressing one of its
simplest parts. We are planing on using more informal and varied corpora in
future tasks as we make progress on these initial issues.
The workshop aims to be a forum where researchers will have a chance to
compare their methods, systems and results.
Important dates
- *March **1*: Registration opened
- *April 17*: Release of the development-set
- *May **12*: Registration deadline
- *May 19*: Release of the test-set
- *May 21*: Result submission deadline
- *May 22-June 12*: Manual evaluation. Publication of results
- *July 3*: Short paper submission deadline
- *July 31*: Papers? camera ready version
- *September **14 *or* 15*: Workshop
Organizing CommitteeI?aki Alegria (UPV/EHU)
Nora Aranberri (UPV/EHU)
Cristina Espa?a-Bonet (UPC)
Pablo Gamallo (USC)
Eva Mart?nez (UPC)
Hugo Oliveira (Universidade de Coimbra)
I?aki San Vicente (Elhuyar)
Antonio Toral (DCU, Dublin)
Arkaitz Zubiaga (University of Warwick)
Proceedings
The papers of the workshop will be published In the proceedings of ?XXXI
Congreso de la Sociedad Espa?ola de Procesamiento de lenguaje natural?.
Proceedings of the workshop will be also published in the CEUR Workshop
Proceedings digital publication service. Additional information
http://komunitatea.elhuyar.org/tweetmt
-------------- next part --------------
An HTML attachment was scrubbed...
URL: http://mailman.mit.edu/mailman/private/moses-support/attachments/20150320/0f15f909/attachment-0001.htm
------------------------------
Message: 2
Date: Fri, 20 Mar 2015 13:19:02 +0100
From: Marcin Junczys-Dowmunt <junczys@amu.edu.pl>
Subject: [Moses-support] Chinese segmentation/tokenization
To: Moses Support <moses-support@mit.edu>
Message-ID: <e4d171cb90994cb853a9965facaebc37@amu.edu.pl>
Content-Type: text/plain; charset="us-ascii"
Hi,
questions appear from time to time on the list concerning Chinese
segmentation/tokenization. I saw Barry mention Lingpipe and other tools.
Is there a favourite tool you guys prefer to use over others?
Thanks,
Marcin
-------------- next part --------------
An HTML attachment was scrubbed...
URL: http://mailman.mit.edu/mailman/private/moses-support/attachments/20150320/64f7c762/attachment-0001.htm
------------------------------
Message: 3
Date: Fri, 20 Mar 2015 12:20:28 +0000
From: Rohit Gupta <enggrohitgupta@gmail.com>
Subject: [Moses-support] CALL FOR BOOK PROPOSALS
To: mt-list@eamt.org, moses-support@mit.edu
Message-ID:
<CAB-CSF-CD-ahwtFtVgXsV12uxU8aUU5g2449wGbd055ufa2eKA@mail.gmail.com>
Content-Type: text/plain; charset="utf-8"
[apologies for cross-posting]
****************************************
* CALL FOR BOOK PROPOSALS
****************************************
John Benjamins' NATURAL LANGUAGE PROCESSING Book Series invites new book
proposals to respond to the growing demand for Natural Language processing
(NLP) literature. Three general types of books are considered for
publication:
----------------------
MONOGRAPHS
----------------------
- original, leading and cutting-edge research (the monograph could be based
on an outstanding PhD thesis)
- surveys of the state of the art in specific NLP tasks or applications
----------------------
COLLECTIONS
----------------------
- books focusing on a particular NLP area (e.g. emerging from successful
NLP workshops or as a result of editors? calls for papers)
- books which include papers covering a wide range of topics (e.g. emerging
from competitive NLP conferences or as a result of proposals for books of
the type "Reading In NLP")
-------------------------
COURSE BOOKS
-------------------------
- general NLP course books
- books on a particular key area of NLP (e.g. Speech Processing,
Computational Syntax/Parsing)
Authors are encouraged to append supplementary materials such as
demonstration programs, NLP software, corpora and so on if applicable, and
to indicate websites and computational language resources where
appropriate. This call invites proposals from potential authors of the
types of books described above. Proposals on any topic related to Natural
Language Processing are welcome.
Interested authors should submit proposals by email (plain text or pdf
files) to the series editor:
Prof. Dr. Ruslan Mitkov
Email R.Mitkov@wlv.ac.uk
with a copy to Emma Franklin (emma.franklin@wlv.ac.uk), the series
editorial assistant.
The proposals should include an outline of the book (1-2 pages), a
preliminary table of contents, the target readership, related publications,
how the book will differ from other similar books in the area (if
applicable), time-scale and information about the prospective author
(relevant experience in the field, publications etc.).
Each proposal will be reviewed by members of the advisory board or
additional reviewers.
For more information on the series, visit:
https://benjamins.com/#catalog/books/nlp/main
----------------
*Rohit Gupta*
*Marie Curie Early Stage Researcher, EXPERT Project*Research Group in
Computational Linguistics
Research Institute of Information and Language Processing
University of Wolverhampton
-------------- next part --------------
An HTML attachment was scrubbed...
URL: http://mailman.mit.edu/mailman/private/moses-support/attachments/20150320/bb9e4ddd/attachment-0001.htm
------------------------------
Message: 4
Date: Fri, 20 Mar 2015 13:32:24 +0000
From: Hieu Hoang <hieuhoang@gmail.com>
Subject: Re: [Moses-support] Problem
To: qinmaoyuan <qinmaoyuan@hotmail.com>, "." <moses-support@mit.edu>,
Ulrich Germann <ugermann@inf.ed.ac.uk>
Message-ID: <550C2168.7080101@gmail.com>
Content-Type: text/plain; charset="utf-8"
There's a problem with compiling with xmlrpc-c at the moment.
https://github.com/moses-smt/mosesdecoder/issues/99
It's being looked at, but in the meantime, try compiling without xmlrpc-c
On 20/03/2015 03:31, qinmaoyuan wrote:
> hi everyone
>
> I am fresh for mosesdecoder. Now i am confused by some problem.
>
> 1.I installed Ubuntu 14.04.02 kylin by using VMware station and some
> software below:
>
> g++
> git
> subversion
> automake
> libtool
> zlib1g-dev
> libboost-all-dev
> libbz2-dev
> liblzma-dev
> python-dev
> libtcmalloc-minimal4
>
> 2.And then i install boost_1_56_0.tar.gz. I downloaded by myself and
> ziped to my directory and installed it successfully.
>
> commands below:
> cd boost_1_55_0/
> ./bootstrap.sh
> ./b2 -j2 --prefix=$PWD --libdir=$PWD/lib64 --layout=tagged link=static
> threading=multi,single install || echo FAILURE
>
> 3.Install xmlrpc-c
> I downloaded this instead of using apt-get.
> commands :
>
> wget http://svn.code.sf.net/p/xmlrpc-c/code
> REPOS=http://svn.code.sf.net/p/xmlrpc-c/code/stable
> svn checkout $REPOS xmlrpc-c
> ./configure --prefix=/usr/local/lib/xml-rpc
> make
> make install
>
> 4.install mosesdecoder
>
> i used git to download mosesdecoder from github and the code is below:
>
> git clone https://github.com/moses-smt/mosesdecoder.git
>
> cd /usr/local/lib/mosesdecoder/
>
> ./bjam --with-xmlrpc-c=/usr/local/lib/xml-rpc
>
> here, system always reported build error.
>
> control platform output :
>
>
>
>
>
>
>
>
> please help me
>
>
> kings regards
>
> Qin
>
>
>
>
>
>
>
>
>
>
>
>
> _______________________________________________
> Moses-support mailing list
> Moses-support@mit.edu
> http://mailman.mit.edu/mailman/listinfo/moses-support
-------------- next part --------------
An HTML attachment was scrubbed...
URL: http://mailman.mit.edu/mailman/private/moses-support/attachments/20150320/dc7e1c77/attachment.htm
------------------------------
_______________________________________________
Moses-support mailing list
Moses-support@mit.edu
http://mailman.mit.edu/mailman/listinfo/moses-support
End of Moses-support Digest, Vol 101, Issue 58
**********************************************
Subscribe to:
Post Comments (Atom)
0 Response to "Moses-support Digest, Vol 101, Issue 58"
Post a Comment