"Buenos dias", "buenas noches" -- this was the first words in a foreign language I heard in my life, as a three-year old boy growing up in developing post-war Western Germany, where the first gastarbeiters had arrived from Spain. Fascinated by the strange sounds, I tried to get to know some more languages, the only opportunity being TV courses of English and French -- there was no foreign language education for pre-teen school children in Germany yet in those days. Read more
To find some answers Tim Machan explores the language's present and past, and looks ahead to its futures among the one and a half billion people who speak it. His search is fascinating and important, for definitions of English have influenced education and law in many countries and helped shape the identities of those who live in them.
This volume provides a new perspective on the evolution of the special language of medicine, based on the electronic corpus of Early Modern English Medical Texts, containing over two million words of medical writing from 1500 to 1700.
'Named Entities' provides critical information for many NLP applications.
Named Entity recognition and classification (NERC) in text is recognized as
one of the important sub-tasks of Information Extraction (IE). The seven
papers in this volume cover various interesting and informative aspects of
NERC research. Nadeau & Sekine provide an extensive survey of past NERC
technologies, which should be a very useful resource for new researchers in
this field. Smith & Osborne describe a machine learning model which tries
to solve the over-fitting problem. Mazur & Dale tackle a common problem
of NE and conjunction; as conjunctions are often a part of NEs or appear
close to NEs, this is an important practical problem. A further three
papers describe analyses and implementations of NERC for different
languages: Spanish (Galicia-Haro & Gelbukh), Bengali (Ekbal, Naskar &;
Bandyopadhyay), and Serbian (Vitas, Krstev & Maurel). Finally, Steinberger
& Pouliquen report on a real WEB application where multilingual NERC
technology is used to identify occurrences of people, locations and
organizations in newspapers in different languages.
The contributions to this volume were previously published as 'Lingvisticae
Investigationes' 30:1 (2007).