Other Workshops and Events (2008)
Volumes
- Software Engineering, Testing, and Quality Assurance for Natural Language Processing 13 papers
- Proceedings of the ACL-08: HLT Workshop on Mobile Language Processing 8 papers
- Proceedings of the Workshop on Parsing German 9 papers
- Coling 2008: Proceedings of the workshop on Human Judgements in Computational Linguistics 10 papers
- Coling 2008: Proceedings of the workshop on Cross-Framework and Cross-Domain Parser Evaluation 9 papers
- Coling 2008: Proceedings of the workshop on Speech Processing for Safety Critical Translation and Pervasive Applications 12 papers
- Coling 2008: Proceedings of the workshop on Knowledge and Reasoning for Answering Questions 7 papers
- Coling 2008: Proceedings of the 2nd workshop on Information Retrieval for Question Answering 11 papers
- Proceedings of the IJCNLP-08 Workshop on NLP for Less Privileged Languages 21 papers
- Proceedings of the Sixth SIGHAN Workshop on Chinese Language Processing 35 papers
- Proceedings of the IJCNLP-08 Workshop on Named Entity Recognition for South and South East Asian Languages 16 papers
- Proceedings of the 2nd workshop on Cross Lingual Information Access (CLIA) Addressing the Information Need of Multilingual Societies 16 papers
- Proceedings of the 6th Workshop on Asian Language Resources 23 papers
- Proceedings of the Workshop on Technologies and Corpora for Asia-Pacific Speech Translation (TCAST) 7 papers
- Proceedings of the 9th SIGdial Workshop on Discourse and Dialogue SIGDIAL 31 papers
- Proceedings of the Third Workshop on Issues in Teaching Computational Linguistics TeachingNLP 17 papers
- Proceedings of the Third Workshop on Statistical Machine Translation WMT 37 papers
- Proceedings of the ACL-08: HLT Second Workshop on Syntax and Structure in Statistical Translation (SSST-2) SSST 12 papers
- Proceedings of the Workshop on Current Trends in Biomedical Natural Language Processing BioNLP 29 papers
- Proceedings of the Tenth Meeting of ACL Special Interest Group on Computational Morphology and Phonology SIGMORPHON 9 papers
- Proceedings of the Third Workshop on Innovative Use of NLP for Building Educational Applications BEA 14 papers
- Coling 2008: Proceedings of the workshop Multi-source Multilingual Information Extraction and Summarization MultiLing 10 papers
- Coling 2008: Proceedings of the workshop on Grammar Engineering Across Frameworks GEAF 9 papers
- Coling 2008: Proceedings of the Workshop on Cognitive Aspects of the Lexicon (COGALEX 2008) CogALex 15 papers
- Coling 2008: Proceedings of the 3rd Textgraphs workshop on Graph-based Algorithms for Natural Language Processing TextGraphs 10 papers
- Proceedings of the Ninth International Workshop on Tree Adjoining Grammar and Related Frameworks (TAG+9) TAG+ 23 papers
- Proceedings of the 5th International Workshop on Spoken Language Translation: Plenaries IWSLT 1 paper
- Proceedings of the 5th International Workshop on Spoken Language Translation: Evaluation Campaign IWSLT 20 papers
- Proceedings of the 5th International Workshop on Spoken Language Translation: Papers IWSLT 8 papers
- Proceedings of the 4th Web as Corpus Workshop WAC 10 papers
- Proceedings of the Australasian Language Technology Association Workshop 2008 ALTA 22 papers
- CoNLL 2008: Proceedings of the Twelfth Conference on Computational Natural Language Learning CoNLL 41 papers
- Proceedings of the Fifth International Natural Language Generation Conference INLG 42 papers
- Semantics in Text Processing. STEP 2008 Conference Proceedings STEP 33 papers
up
Software Engineering, Testing, and Quality Assurance for Natural Language Processing
Software Engineering, Testing, and Quality Assurance for Natural Language Processing
K. Bretonnel Cohen | Bob Carpenter
K. Bretonnel Cohen | Bob Carpenter
Increasing Maintainability of NLP Evaluation Modules Through Declarative Implementations
Terry Heinze | Marc Light
Terry Heinze | Marc Light
zymake: A Computational Workflow System for Machine Learning and Natural Language Processing
Eric Breck
Eric Breck
Evaluating the Effects of Treebank Size in a Practical Application for Parsing
Kenji Sagae | Yusuke Miyao | Rune Saetre | Jun’ichi Tsujii
Kenji Sagae | Yusuke Miyao | Rune Saetre | Jun’ichi Tsujii
Adapting Naturally Occurring Test Suites for Evaluation of Clinical Question Answering
Dina Demner-Fushman
Dina Demner-Fushman
Software Testing and the Naturally Occurring Data Assumption in Natural Language Processing
K. Bretonnel Cohen | William A. Baumgartner Jr. | Lawrence Hunter
K. Bretonnel Cohen | William A. Baumgartner Jr. | Lawrence Hunter
Building a BioWordNet Using WordNet Data Structures and WordNet’s Software Infrastructure–A Failure Story
Michael Poprat | Elena Beisswanger | Udo Hahn
Michael Poprat | Elena Beisswanger | Udo Hahn
Fast, Scalable and Reliable Generation of Controlled Natural Language
David Hardcastle | Richard Power
David Hardcastle | Richard Power
up
Proceedings of the ACL-08: HLT Workshop on Mobile Language Processing
A Multimodal Home Entertainment Interface via a Mobile Device
Alexander Gruenstein | Bo-June Paul Hsu | James Glass | Stephanie Seneff | Lee Hetherington | Scott Cyphers | Ibrahim Badr | Chao Wang | Sean Liu
Alexander Gruenstein | Bo-June Paul Hsu | James Glass | Stephanie Seneff | Lee Hetherington | Scott Cyphers | Ibrahim Badr | Chao Wang | Sean Liu
A Wearable Headset Speech-to-Speech Translation System
Kriste Krstovski | Michael Decerbo | Rohit Prasad | David Stallard | Shirin Saleem | Premkumar Natarajan
Kriste Krstovski | Michael Decerbo | Rohit Prasad | David Stallard | Shirin Saleem | Premkumar Natarajan
Information extraction using finite state automata and syllable n-grams in a mobile environment
Choong-Nyoung Seon | Harksoo Kim | Jungyun Seo
Choong-Nyoung Seon | Harksoo Kim | Jungyun Seo
up
Proceedings of the Workshop on Parsing German
Part-of-Speech Tagging with a Symbolic Full Parser: Using the TIGER Treebank to Evaluate Fips
Yves Scherrer
Yves Scherrer
Revisiting the Impact of Different Annotation Schemes on PCFG Parsing: A Grammatical Dependency Evaluation
Adriane Boyd | Detmar Meurers
Adriane Boyd | Detmar Meurers
Parsing Three German Treebanks: Lexicalized and Unlexicalized Baselines
Anna Rafferty | Christopher D. Manning
Anna Rafferty | Christopher D. Manning
up
Coling 2008: Proceedings of the workshop on Human Judgements in Computational Linguistics
Coling 2008: Proceedings of the workshop on Human Judgements in Computational Linguistics
Ron Artstein | Gemma Boleda | Frank Keller | Sabine Schulte im Walde
Ron Artstein | Gemma Boleda | Frank Keller | Sabine Schulte im Walde
Invited Talk: The Relevance of a Cognitive Model of the Mental Lexicon to Automatic Word Sense Disambiguation
Martha Palmer | Susan Brown
Martha Palmer | Susan Brown
Human Judgement as a Parameter in Evaluation Campaigns
Jean-Baptiste Berthelin | Cyril Grouin | Martine Hurault-Plantet | Patrick Paroubek
Jean-Baptiste Berthelin | Cyril Grouin | Martine Hurault-Plantet | Patrick Paroubek
Native Judgments of Non-Native Usage: Experiments in Preposition Error Detection
Joel Tetreault | Martin Chodorow
Joel Tetreault | Martin Chodorow
up
Coling 2008: Proceedings of the workshop on Cross-Framework and Cross-Domain Parser Evaluation
Coling 2008: Proceedings of the workshop on Cross-Framework and Cross-Domain Parser Evaluation
Johan Bos | Edward Briscoe | Aoife Cahill | John Carroll | Stephen Clark | Ann Copestake | Dan Flickinger | Josef van Genabith | Julia Hockenmaier | Aravind Joshi | Ronald Kaplan | Tracy Holloway King | Sandra Kuebler | Dekang Lin | Jan Tore Lønning | Christopher Manning | Yusuke Miyao | Joakim Nivre | Stephan Oepen | Kenji Sagae | Nianwen Xue | Yi Zhang
Johan Bos | Edward Briscoe | Aoife Cahill | John Carroll | Stephen Clark | Ann Copestake | Dan Flickinger | Josef van Genabith | Julia Hockenmaier | Aravind Joshi | Ronald Kaplan | Tracy Holloway King | Sandra Kuebler | Dekang Lin | Jan Tore Lønning | Christopher Manning | Yusuke Miyao | Joakim Nivre | Stephan Oepen | Kenji Sagae | Nianwen Xue | Yi Zhang
Exploring an Auxiliary Distribution Based Approach to Domain Adaptation of a Syntactic Disambiguation Model
Barbara Plank | Gertjan van Noord
Barbara Plank | Gertjan van Noord
Parser Evaluation Across Frameworks without Format Conversion
Wai Lok Tam | Yo Sato | Yusuke Miyao | Junichi Tsujii
Wai Lok Tam | Yo Sato | Yusuke Miyao | Junichi Tsujii
up
Coling 2008: Proceedings of the workshop on Speech Processing for Safety Critical Translation and Pervasive Applications
Coling 2008: Proceedings of the workshop on Speech Processing for Safety Critical Translation and Pervasive Applications
Pierrette Bouillon | Farzad Ehsani | Robert Frederking | Michael McTear | Manny Rayner
Pierrette Bouillon | Farzad Ehsani | Robert Frederking | Michael McTear | Manny Rayner
Mitigation of Data Sparsity in Classifier-Based Translation
Emil Ettelaie | Panayiotis G. Georgiou | Shrikanth S. Narayanan
Emil Ettelaie | Panayiotis G. Georgiou | Shrikanth S. Narayanan
An Integrated Dialog Simulation Technique for Evaluating Spoken Dialog Systems
Sangkeun Jung | Cheongjae Lee | Kyungduk Kim | Gary Geunbae Lee
Sangkeun Jung | Cheongjae Lee | Kyungduk Kim | Gary Geunbae Lee
Economical Global Access to a VoiceXML Gateway Using Open Source Technologies
Kulwinder Singh | Dong-Won Park
Kulwinder Singh | Dong-Won Park
Interoperability and Knowledge Representation in Distributed Health and Fitness Companion Dialogue System
Jaakko Hakulinen | Markku Turunen
Jaakko Hakulinen | Markku Turunen
The 2008 MedSLT System
Manny Rayner | Pierrette Bouillon | Jane Brotanek | Glenn Flores | Sonia Halimi | Beth Ann Hockey | Hitoshi Isahara | Kyoko Kanzaki | Elisabeth Kron | Yukie Nakao | Marianne Santaholma | Marianne Starlander | Nikos Tsourakis
Manny Rayner | Pierrette Bouillon | Jane Brotanek | Glenn Flores | Sonia Halimi | Beth Ann Hockey | Hitoshi Isahara | Kyoko Kanzaki | Elisabeth Kron | Yukie Nakao | Marianne Santaholma | Marianne Starlander | Nikos Tsourakis
Language Understanding in Maryland Virtual Patient
Sergei Nirenburg | Stephen Beale | Marjorie McShane | Bruce Jarrell | George Fantry
Sergei Nirenburg | Stephen Beale | Marjorie McShane | Bruce Jarrell | George Fantry
Rapid Portability among Domains in an Interactive Spoken Language Translation System
Mark Seligman | Mike Dillinger
Mark Seligman | Mike Dillinger
Speech Translation for Triage of Emergency Phonecalls in Minority Languages
Udhyakumar Nallasamy | Alan Black | Tanja Schultz | Robert Frederking | Jerry Weltman
Udhyakumar Nallasamy | Alan Black | Tanja Schultz | Robert Frederking | Jerry Weltman
up
Coling 2008: Proceedings of the workshop on Knowledge and Reasoning for Answering Questions
Coling 2008: Proceedings of the workshop on Knowledge and Reasoning for Answering Questions
Marie-Francine Moens | Patrick Saint-Dizier
Marie-Francine Moens | Patrick Saint-Dizier
Semantic Chunk Annotation for complex questions using Conditional Random Field
Shixi Fan | Yaoyun Zhang | Wing W. Y. Ng | Xuan Wang | Xiaolong Wang
Shixi Fan | Yaoyun Zhang | Wing W. Y. Ng | Xuan Wang | Xiaolong Wang
up
Coling 2008: Proceedings of the 2nd workshop on Information Retrieval for Question Answering
Coling 2008: Proceedings of the 2nd workshop on Information Retrieval for Question Answering
Mark A. Greenwood
Mark A. Greenwood
Improving Text Retrieval Precision and Answer Accuracy in Question Answering Systems
Matthew Bilotti | Eric Nyberg
Matthew Bilotti | Eric Nyberg
Exact Phrases in Information Retrieval for Question Answering
Svetlana Stoyanchev | Young Chol Song | William Lahti
Svetlana Stoyanchev | Young Chol Song | William Lahti
Simple is Best: Experiments with Different Document Segmentation Strategies for Passage Retrieval
Jörg Tiedemann | Jori Mur
Jörg Tiedemann | Jori Mur
A Data Driven Approach to Query Expansion in Question Answering
Leon Derczynski | Jun Wang | Robert Gaizauskas | Mark A. Greenwood
Leon Derczynski | Jun Wang | Robert Gaizauskas | Mark A. Greenwood
Using Lexico-Semantic Information for Query Expansion in Passage Retrieval for Question Answering
Lonneke van der Plas | Jörg Tiedemann
Lonneke van der Plas | Jörg Tiedemann
up
Proceedings of the IJCNLP-08 Workshop on NLP for Less Privileged Languages
Natural Language Processing for Less Privileged Languages: Where do we come from? Where are we going?
Anil Kumar Singh
Anil Kumar Singh
KUI: an ubiquitous tool for collective intelligence development
Thatsanee Charoenporn | Virach Sornlertlamvanich | Hitoshi Isahara | Kergrit Robkop
Thatsanee Charoenporn | Virach Sornlertlamvanich | Hitoshi Isahara | Kergrit Robkop
Prototype Machine Translation System From Text-To-Indian Sign Language
Tirthankar Dasgupta | Sandipan Dandpat | Anupam Basu
Tirthankar Dasgupta | Sandipan Dandpat | Anupam Basu
SriShell Primo: A Predictive Sinhala Text Input System
Sandeva Goonetilleke | Yoshihiko Hayashi | Yuichi Itoh | Fumio Kishino
Sandeva Goonetilleke | Yoshihiko Hayashi | Yuichi Itoh | Fumio Kishino
Strategies for sustainable MT for Basque: incremental design, reusability, standardization and open-source
I. Alegria | X. Arregi | X. Artola | A. Diaz de Ilarraza | G. Labaka | M. Lersundi | A. Mayor | K. Sarasola
I. Alegria | X. Arregi | X. Artola | A. Diaz de Ilarraza | G. Labaka | M. Lersundi | A. Mayor | K. Sarasola
Design of a Rule-based Stemmer for Natural Language Text in Bengali
Sandipan Sarkar | Sivaji Bandyopadhyay
Sandipan Sarkar | Sivaji Bandyopadhyay
up
Proceedings of the Sixth SIGHAN Workshop on Chinese Language Processing
An Example-based Decoder for Spoken Language Machine Translation
Zhou-Jun Li | Wen-Han Chao | Yue-Xin Chen
Zhou-Jun Li | Wen-Han Chao | Yue-Xin Chen
Automatic Extraction of English-Chinese Transliteration Pairs using Dynamic Window and Tokenizer
Chengguo Jin | Seung-Hoon Na | Dong-Il Kim | Jong-Hyeok Lee
Chengguo Jin | Seung-Hoon Na | Dong-Il Kim | Jong-Hyeok Lee
Mining Transliterations from Web Query Results: An Incremental Approach
Jin-Shea Kuo | Haizhou Li | Chih-Lung Lin
Jin-Shea Kuo | Haizhou Li | Chih-Lung Lin
Use of Event Types for Temporal Relation Identification in Chinese Text
Yuchang Cheng | Masayuki Asahara | Yuji Matsumoto
Yuchang Cheng | Masayuki Asahara | Yuji Matsumoto
Analyzing Chinese Synthetic Words with Tree-based Information and a Survey on Chinese Morphologically Derived Words
Jia Lu | Masayuki Asahara | Yuji Matsumoto
Jia Lu | Masayuki Asahara | Yuji Matsumoto
Which Performs Better on In-Vocabulary Word Segmentation: Based on Word or Character?
Zhenxing Wang | Changning Huang | Jingbo Zhu
Zhenxing Wang | Changning Huang | Jingbo Zhu
The Fourth International Chinese Language Processing Bakeoff: Chinese Word Segmentation, Named Entity Recognition and Chinese POS Tagging
Guangjin Jin | Xiao Chen
Guangjin Jin | Xiao Chen
NOKIA Research Center Beijing Chinese Word Segmentation System for the SIGHAN Bakeoff 2007
Jiang Li | Rile Hu | Guohua Zhang | Yuezhong Tang | Zhanjiang Song | Xia Wang
Jiang Li | Rile Hu | Guohua Zhang | Yuezhong Tang | Zhanjiang Song | Xia Wang
Chinese Word Segmentation and Named Entity Recognition Based on Conditional Random Fields
Xinnian Mao | Yuan Dong | Saike He | Sencheng Bao | Haila Wang
Xinnian Mao | Yuan Dong | Saike He | Sencheng Bao | Haila Wang
The Character-based CRF Segmenter of MSRA&NEU for the 4th Bakeoff
Zhenxing Wang | Changning Huang | Jingbo Zhu
Zhenxing Wang | Changning Huang | Jingbo Zhu
Chinese NER Using CRFs and Logic for the Fourth SIGHAN Bakeoff
Xiaofeng Yu | Wai Lam | Shing-Kit Chan | Yiu Kei Wu | Bo Chen
Xiaofeng Yu | Wai Lam | Shing-Kit Chan | Yiu Kei Wu | Bo Chen
Unsupervised Segmentation Helps Supervised Learning of Character Tagging for Word Segmentation and Named Entity Recognition
Hai Zhao | Chunyu Kit
Hai Zhao | Chunyu Kit
Nanjing Normal University Segmenter for the Fourth SIGHAN Bakeoff
Xiaohe Chen | Bin Li | Junzhi Lu | Hongdong Nian | Xuri Tang
Xiaohe Chen | Bin Li | Junzhi Lu | Hongdong Nian | Xuri Tang
Two Step Chinese Named Entity Recognition Based on Conditional Random Fields Models
Yuanyong Feng | Ruihong Huang | Le Sun
Yuanyong Feng | Ruihong Huang | Le Sun
HMM and CRF Based Hybrid Model for Chinese Lexical Analysis
Degen Huang | Xiao Sun | Shidou Jiao | Lishuang Li | Zhuoye Ding | Ru Wan
Degen Huang | Xiao Sun | Shidou Jiao | Lishuang Li | Zhuoye Ding | Ru Wan
Training a Perceptron with Global and Local Features for Chinese Word Segmentation
Dong Song | Anoop Sarkar
Dong Song | Anoop Sarkar
A Study of Chinese Lexical Analysis Based on Discriminative Models
Guang-Lu Sun | Cheng-Jie Sun | Ke Sun | Xiao-Long Wang
Guang-Lu Sun | Cheng-Jie Sun | Ke Sun | Xiao-Long Wang
An Improved CRF based Chinese Language Processing System for SIGHAN Bakeoff 2007
Xihong Wu | Xiaojun Lin | Xinhao Wang | Chunyao Wu | Yaozhong Zhang | Dianhai Yu
Xihong Wu | Xiaojun Lin | Xinhao Wang | Chunyao Wu | Yaozhong Zhang | Dianhai Yu
Description of the NCU Chinese Word Segmentation and Part-of-Speech Tagging for SIGHAN Bakeoff 2007
Yu-Chieh Wu | Jie-Chi Yang | Yue-Shi Lee
Yu-Chieh Wu | Jie-Chi Yang | Yue-Shi Lee
CRF-based Hybrid Model for Word Segmentation, NER and even POS Tagging
Zhiting Xu | Xian Qian | Yuejie Zhang | Yaqian Zhou
Zhiting Xu | Xian Qian | Yuejie Zhang | Yaqian Zhou
CRFs-Based Named Entity Recognition Incorporated with Heuristic Entity List Searching
Fan Yang | Jun Zhao | Bo Zou
Fan Yang | Jun Zhao | Bo Zou
A Chinese Word Segmentation System Based on Cascade Model
Jianfeng Zhang | Jiaheng Zheng | Hu Zhang | Hongye Tan
Jianfeng Zhang | Jiaheng Zheng | Hu Zhang | Hongye Tan
up
Proceedings of the IJCNLP-08 Workshop on Named Entity Recognition for South and South East Asian Languages
A Hybrid Named Entity Recognition System for South and South East Asian Languages
Sujan Kumar Saha | Sanjay Chatterji | Sandipan Dandapat | Sudeshna Sarkar | Pabitra Mitra
Sujan Kumar Saha | Sanjay Chatterji | Sandipan Dandapat | Sudeshna Sarkar | Pabitra Mitra
Aggregating Machine Learning and Rule Based Heuristics for Named Entity Recognition
Karthik Gali | Harshit Surana | Ashwini Vaidya | Praneeth Shishtla | Dipti Misra Sharma
Karthik Gali | Harshit Surana | Ashwini Vaidya | Praneeth Shishtla | Dipti Misra Sharma
Language Independent Named Entity Recognition in Indian Languages
Asif Ekbal | Rejwanul Haque | Amitava Das | Venkateswarlu Poka | Sivaji Bandyopadhyay
Asif Ekbal | Rejwanul Haque | Amitava Das | Venkateswarlu Poka | Sivaji Bandyopadhyay
Domain Focused Named Entity Recognizer for Tamil Using Conditional Random Fields
Vijayakrishna R | Sobha L
Vijayakrishna R | Sobha L
A Character n-gram Based Approach for Improved Recall in Indian Language NER
Praneeth M Shishtla | Prasad Pingali | Vasudeva Varma
Praneeth M Shishtla | Prasad Pingali | Vasudeva Varma
An Experiment on Automatic Detection of Named Entities in Bangla
Bidyut Baran Chaudhuri | Suvankar Bhattacharya
Bidyut Baran Chaudhuri | Suvankar Bhattacharya
Hybrid Named Entity Recognition System for South and South East Asian Languages
Praveen P | Ravi Kiran V
Praveen P | Ravi Kiran V
up
Proceedings of the 2nd workshop on Cross Lingual Information Access (CLIA) Addressing the Information Need of Multilingual Societies
The Effects of Language Relatedness on Multilingual Information Retrieval: A Case Study With Indo-European and Semitic Languages
Peter Chew | Ahmed Abdelali
Peter Chew | Ahmed Abdelali
Some Experiments in Mining Named Entity Transliteration Pairs from Comparable Corpora
K Saravanan | A Kumaran
K Saravanan | A Kumaran
Domain-Specific Query Translation for Multilingual Information Access using Machine Translation Augmented With Dictionaries Mined from Wikipedia
Gareth Jones | Fabio Fantino | Eamonn Newman | Ying Zhang
Gareth Jones | Fabio Fantino | Eamonn Newman | Ying Zhang
Statistical Transliteration for Cross Language Information Retrieval using HMM alignment model and CRF
Prasad Pingali | Surya Ganesh | Sreeharsha Yella | Vasudeva Varma
Prasad Pingali | Surya Ganesh | Sreeharsha Yella | Vasudeva Varma
Script Independent Word Spotting in Multilingual Documents
Anurag Bhardwaj | Damien Jose | Venu Govindaraju
Anurag Bhardwaj | Damien Jose | Venu Govindaraju
A Document Graph Based Query Focused Multi-Document Summarizer
Sibabrata Paladhi | Sivaji Bandyopadhyay
Sibabrata Paladhi | Sivaji Bandyopadhyay
Hindi and Marathi to English Cross Language Information Retrieval
Manoj Kumar Chinnakotla | Sagar Ranadive | Om P. Damani | Pushpak Bhattacharyya
Manoj Kumar Chinnakotla | Sagar Ranadive | Om P. Damani | Pushpak Bhattacharyya
Bengali and Hindi to English CLIR Evaluation
Debasis Mandal | Sandipan Dandapat | Mayank Gupta | Pratyush Banerjee | Sudeshna Sarkar
Debasis Mandal | Sandipan Dandapat | Mayank Gupta | Pratyush Banerjee | Sudeshna Sarkar
up
Proceedings of the 6th Workshop on Asian Language Resources
Development of Bengali Named Entity Tagged Corpus and its Use in NER Systems
Asif Ekbal | Sivaji Bandyopadhyay
Asif Ekbal | Sivaji Bandyopadhyay
Gazetteer Preparation for Named Entity Recognition in Indian Languages
Sujan Kumar Saha | Sudeshna Sarkar | Pabitra Mitra
Sujan Kumar Saha | Sudeshna Sarkar | Pabitra Mitra
Technical Terminology in Asian Languages: Different Approaches to Adopting Engineering Terms
Makiko Matsuda | Tomoe Takahashi | Hiroki Goto | Yoshikazu Hayase | Robin Lee Nagano | Yoshiki Mikami
Makiko Matsuda | Tomoe Takahashi | Hiroki Goto | Yoshikazu Hayase | Robin Lee Nagano | Yoshiki Mikami
The Link Structure of Language Communities and its Implication for Language-specific Crawling
Rizza Caminero | Yoshiki Mikami
Rizza Caminero | Yoshiki Mikami
A Multilingual Multimedia Indian Sign Language Dictionary Tool
Tirthankar Dasgupta | Sambit Shukla | Sandeep Kumar | Synny Diwakar | Anupam Basu
Tirthankar Dasgupta | Sambit Shukla | Sandeep Kumar | Synny Diwakar | Anupam Basu
A Discourse Resource for Turkish: Annotating Discourse Connectives in the METU Corpus
Deniz Zeyrek | Bonnie Webber
Deniz Zeyrek | Bonnie Webber
Towards an Annotated Corpus of Discourse Relations in Hindi
Rashmi Prasad | Samar Husain | Dipti Sharma | Aravind Joshi
Rashmi Prasad | Samar Husain | Dipti Sharma | Aravind Joshi
A Semantic Study on Yami Ontology in Traditional Songs
Yin-Sheng Tai | D. Victoria Rau | Meng-Chien Yang
Yin-Sheng Tai | D. Victoria Rau | Meng-Chien Yang
Assessment and Development of POS Tag Set for Telugu
Rama Sree R.J | Uma Maheswara Rao G | Madhu Murthy K.V
Rama Sree R.J | Uma Maheswara Rao G | Madhu Murthy K.V
Designing a Common POS-Tagset Framework for Indian Languages
Sankaran Baskaran | Kalika Bali | Tanmoy Bhattacharya | Pushpak Bhattacharyya | Girish Nath Jha | Rajendran S | Saravanan K | Sobha L | Subbarao K V.
Sankaran Baskaran | Kalika Bali | Tanmoy Bhattacharya | Pushpak Bhattacharyya | Girish Nath Jha | Rajendran S | Saravanan K | Sobha L | Subbarao K V.
Confirmed Language Resource for Answering How Type Questions Developed by Using Mails Posted to a Mailing List
Ryo Nishimura | Yasuhiko Watanabe | Yoshihiro Okada
Ryo Nishimura | Yasuhiko Watanabe | Yoshihiro Okada
A Basic Framework to Build a Test Collection for the Vietnamese Text Catergorization
Viet Hoang-Anh | Thu Dinh-Thi-Phuong | Thang Huynh-Quyet
Viet Hoang-Anh | Thu Dinh-Thi-Phuong | Thang Huynh-Quyet
up
Proceedings of the Workshop on Technologies and Corpora for Asia-Pacific Speech Translation (TCAST)
Transformation-based Sentence Splitting method for Statistical Machine Translation
Jonghoon Lee | Donghyeon Lee | Gary Geunbae Lee
Jonghoon Lee | Donghyeon Lee | Gary Geunbae Lee
Speech-to-Speech Translation Activities in Thailand
Chai Wutiwiwatchai | Thepchai Supnithi | Krit Kosawat
Chai Wutiwiwatchai | Thepchai Supnithi | Krit Kosawat
Development of Indonesian Large Vocabulary Continuous Speech Recognition System within A-STAR Project
Sakriani Sakti | Eka Kelana | Hammam Riza | Shinsuke Sakai | Konstantin Markov | Satoshi Nakamura
Sakriani Sakti | Eka Kelana | Hammam Riza | Shinsuke Sakai | Konstantin Markov | Satoshi Nakamura
up
Proceedings of the 9th SIGdial Workshop on Discourse and Dialogue
Optimizing Endpointing Thresholds using Dialogue Features in a Spoken Dialogue System
Antoine Raux | Maxine Eskenazi
Antoine Raux | Maxine Eskenazi
Learning N-Best Correction Models from Implicit User Feedback in a Multi-Modal Local Search Application
Dan Bohus | Xiao Li | Patrick Nguyen | Geoffrey Zweig
Dan Bohus | Xiao Li | Patrick Nguyen | Geoffrey Zweig
Reactive Redundancy and Listener Comprehension in Direction-Giving
Rachel Baker | Alastair Gill | Justine Cassell
Rachel Baker | Alastair Gill | Justine Cassell
Rapidly Deploying Grammar-Based Speech Applications with Active Learning and Back-off Grammars
Tim Paek | Sudeep Gandhe | Max Chickering
Tim Paek | Sudeep Gandhe | Max Chickering
Persistent Information State in a Data-Centric Architecture
Sebastian Varges | Giuseppe Riccardi | Silvia Quarteroni
Sebastian Varges | Giuseppe Riccardi | Silvia Quarteroni
What Are Meeting Summaries? An Analysis of Human Extractive Summaries in Meeting Corpus
Fei Liu | Yang Liu
Fei Liu | Yang Liu
A Simple Method for Resolution of Definite Reference in a Shared Visual Context
Alexander Siebert | David Schlangen
Alexander Siebert | David Schlangen
A Framework for Building Conversational Agents Based on a Multi-Expert Model
Mikio Nakano | Kotaro Funakoshi | Yuji Hasegawa | Hiroshi Tsujino
Mikio Nakano | Kotaro Funakoshi | Yuji Hasegawa | Hiroshi Tsujino
From GEMINI to DiaGen: Improving Development of Speech Dialogues for Embedded Systems
Stefan Hamerich
Stefan Hamerich
Quantifying Ellipsis in Dialogue: an index of mutual understanding
Marcus Colman | Arash Eshghi | Pat Healey
Marcus Colman | Arash Eshghi | Pat Healey
Implicit Proposal Filtering in Multi-Party Consensus-Building Conversations
Yasuhiro Katagiri | Yosuke Matsusaka | Yasuharu Den | Mika Enomoto | Masato Ishizaki | Katsuya Takanashi
Yasuhiro Katagiri | Yosuke Matsusaka | Yasuharu Den | Mika Enomoto | Masato Ishizaki | Katsuya Takanashi
Optimal Dialog in Consumer-Rating Systems using POMDP Framework
Zhifei Li | Patrick Nguyen | Geoffrey Zweig
Zhifei Li | Patrick Nguyen | Geoffrey Zweig
Training and Evaluation of the HIS POMDP Dialogue System in Noise
Milica Gašić | Simon Keizer | Francois Mairesse | Jost Schatzmann | Blaise Thomson | Kai Yu | Steve Young
Milica Gašić | Simon Keizer | Francois Mairesse | Jost Schatzmann | Blaise Thomson | Kai Yu | Steve Young
A Frame-Based Probabilistic Framework for Spoken Dialog Management Using Dialog Examples
Kyungduk Kim | Cheongjae Lee | Sangkeun Jung | Gary Geunbae Lee
Kyungduk Kim | Cheongjae Lee | Sangkeun Jung | Gary Geunbae Lee
Speaking More Like You: Lexical, Acoustic/Prosodic, and Discourse Entrainment in Spoken Dialogue Systems
Julia Hirschberg
Julia Hirschberg
Discourse Level Opinion Relations: An Annotation Study
Swapna Somasundaran | Josef Ruppenhofer | Janyce Wiebe
Swapna Somasundaran | Josef Ruppenhofer | Janyce Wiebe
Argumentative Human Computer Dialogue for Automated Persuasion
Pierre Andrews | Suresh Manandhar | Marco De Boni
Pierre Andrews | Suresh Manandhar | Marco De Boni
Modeling Vocal Interaction for Text-Independent Participant Characterization in Multi-Party Conversation
Kornel Laskowski | Mari Ostendorf | Tanja Schultz
Kornel Laskowski | Mari Ostendorf | Tanja Schultz
Modelling and Detecting Decisions in Multi-party Dialogue
Raquel Fernández | Matthew Frampton | Patrick Ehlen | Matthew Purver | Stanley Peters
Raquel Fernández | Matthew Frampton | Patrick Ehlen | Matthew Purver | Stanley Peters
up
Proceedings of the Third Workshop on Issues in Teaching Computational Linguistics
Proceedings of the Third Workshop on Issues in Teaching Computational Linguistics
Martha Palmer | Chris Brew | Fei Xia
Martha Palmer | Chris Brew | Fei Xia
Teaching Computational Linguistics to a Large, Diverse Student Body: Courses, Tools, and Interdepartmental Interaction
Jason Baldridge | Katrin Erk
Jason Baldridge | Katrin Erk
Building a Flexible, Collaborative, Intensive Master’s Program in Computational Linguistics
Emily M. Bender | Fei Xia | Erik Bansleben
Emily M. Bender | Fei Xia | Erik Bansleben
Defining a Core Body of Knowledge for the Introductory Computational Linguistics Curriculum
Steven Bird
Steven Bird
Multidisciplinary Instruction with the Natural Language Toolkit
Steven Bird | Ewan Klein | Edward Loper | Jason Baldridge
Steven Bird | Ewan Klein | Edward Loper | Jason Baldridge
Combining Open-Source with Research to Re-engineer a Hands-on Introductory NLP Course
Nitin Madnani | Bonnie J. Dorr
Nitin Madnani | Bonnie J. Dorr
Zero to Spoken Dialogue System in One Quarter: Teaching Computational Linguistics to Linguists Using Regulus
Beth Ann Hockey | Gwen Christian
Beth Ann Hockey | Gwen Christian
The North American Computational Linguistics Olympiad (NACLO)
Dragomir R. Radev | Lori Levin | Thomas E. Payne
Dragomir R. Radev | Lori Levin | Thomas E. Payne
up
Proceedings of the Third Workshop on Statistical Machine Translation
Proceedings of the Third Workshop on Statistical Machine Translation
Chris Callison-Burch | Philipp Koehn | Christof Monz | Josh Schroeder | Cameron Shaw Fordyce
Chris Callison-Burch | Philipp Koehn | Christof Monz | Josh Schroeder | Cameron Shaw Fordyce
An Empirical Study in Source Word Deletion for Phrase-Based Statistical Machine Translation
Chi-Ho Li | Hailei Zhang | Dongdong Zhang | Mu Li | Ming Zhou
Chi-Ho Li | Hailei Zhang | Dongdong Zhang | Mu Li | Ming Zhou
Regularization and Search for Minimum Error Rate Training
Daniel Cer | Dan Jurafsky | Christopher D. Manning
Daniel Cer | Dan Jurafsky | Christopher D. Manning
Learning Performance of a Machine Translation System: a Statistical and Computational Analysis
Marco Turchi | Tijl De Bie | Nello Cristianini
Marco Turchi | Tijl De Bie | Nello Cristianini
Using Syntax to Improve Word Alignment Precision for Syntax-Based Machine Translation
Victoria Fossum | Kevin Knight | Steven Abney
Victoria Fossum | Kevin Knight | Steven Abney
Using Shallow Syntax Information to Improve Word Alignment and Reordering for SMT
Josep M. Crego | Nizar Habash
Josep M. Crego | Nizar Habash
Further Meta-Evaluation of Machine Translation
Chris Callison-Burch | Cameron Fordyce | Philipp Koehn | Christof Monz | Josh Schroeder
Chris Callison-Burch | Cameron Fordyce | Philipp Koehn | Christof Monz | Josh Schroeder
Limsi’s Statistical Translation Systems for WMT‘08
Daniel Déchelotte | Gilles Adda | Alexandre Allauzen | Hélène Bonneau-Maynard | Olivier Galibert | Jean-Luc Gauvain | Philippe Langlais | François Yvon
Daniel Déchelotte | Gilles Adda | Alexandre Allauzen | Hélène Bonneau-Maynard | Olivier Galibert | Jean-Luc Gauvain | Philippe Langlais | François Yvon
Meteor, M-BLEU and M-TER: Evaluation Metrics for High-Correlation with Human Rankings of Machine Translation Output
Abhaya Agarwal | Alon Lavie
Abhaya Agarwal | Alon Lavie
First Steps towards a General Purpose French/English Statistical Machine Translation System
Holger Schwenk | Jean-Baptiste Fouet | Jean Senellart
Holger Schwenk | Jean-Baptiste Fouet | Jean Senellart
The University of Washington Machine Translation System for ACL WMT 2008
Amittai Axelrod | Mei Yang | Kevin Duh | Katrin Kirchhoff
Amittai Axelrod | Mei Yang | Kevin Duh | Katrin Kirchhoff
The TALP-UPC Ngram-Based Statistical Machine Translation System for ACL-WMT 2008
Maxim Khalilov | Adolfo Hernández H. | Marta R. Costa-jussà | Josep M. Crego | Carlos A. Henríquez Q. | Patrik Lambert | José A. R. Fonollosa | José B. Mariño | Rafael E. Banchs
Maxim Khalilov | Adolfo Hernández H. | Marta R. Costa-jussà | Josep M. Crego | Carlos A. Henríquez Q. | Patrik Lambert | José A. R. Fonollosa | José B. Mariño | Rafael E. Banchs
European Language Translation with Weighted Finite State Transducers: The CUED MT System for the 2008 ACL Workshop on SMT
Graeme Blackwood | Adrià de Gispert | Jamie Brunning | William Byrne
Graeme Blackwood | Adrià de Gispert | Jamie Brunning | William Byrne
Effects of Morphological Analysis in Translation between German and English
Sara Stymne | Maria Holmqvist | Lars Ahrenberg
Sara Stymne | Maria Holmqvist | Lars Ahrenberg
Towards better Machine Translation Quality for the German-English Language Pairs
Philipp Koehn | Abhishek Arun | Hieu Hoang
Philipp Koehn | Abhishek Arun | Hieu Hoang
Phrase-Based and Deep Syntactic English-to-Czech Statistical Machine Translation
Ondřej Bojar | Jan Hajič
Ondřej Bojar | Jan Hajič
Improving English-Spanish Statistical Machine Translation: Experiments in Domain Adaptation, Sentence Paraphrasing, Tokenization, and Recasing
Preslav Nakov
Preslav Nakov
Improving Word Alignment with Language Model Based Confidence Scores
Nguyen Bach | Qin Gao | Stephan Vogel
Nguyen Bach | Qin Gao | Stephan Vogel
Kernel Regression Framework for Machine Translation: UCL System Description for WMT 2008 Shared Translation Task
Zhuoran Wang | John Shawe-Taylor
Zhuoran Wang | John Shawe-Taylor
Using Syntactic Coupling Features for Discriminating Phrase-Based Translations (WMT-08 Shared Translation Task)
Vassilina Nikoulina | Marc Dymetman
Vassilina Nikoulina | Marc Dymetman
Statistical Transfer Systems for French-English and German-English Machine Translation
Greg Hanneman | Edmund Huber | Abhaya Agarwal | Vamshi Ambati | Alok Parlikar | Erik Peterson | Alon Lavie
Greg Hanneman | Edmund Huber | Abhaya Agarwal | Vamshi Ambati | Alok Parlikar | Erik Peterson | Alon Lavie
TectoMT: Highly Modular MT System with Tectogrammatics Used as Transfer Layer
Zdeněk Žabokrtský | Jan Ptáček | Petr Pajas
Zdeněk Žabokrtský | Jan Ptáček | Petr Pajas
Using Moses to Integrate Multiple Rule-Based Machine Translation Engines into a Hybrid System
Andreas Eisele | Christian Federmann | Hervé Saint-Amand | Michael Jellinghaus | Teresa Herrmann | Yu Chen
Andreas Eisele | Christian Federmann | Hervé Saint-Amand | Michael Jellinghaus | Teresa Herrmann | Yu Chen
Incremental Hypothesis Alignment for Building Confusion Networks with Application to Machine Translation System Combination
Antti-Veikko Rosti | Bing Zhang | Spyros Matsoukas | Richard Schwartz
Antti-Veikko Rosti | Bing Zhang | Spyros Matsoukas | Richard Schwartz
Fast, Easy, and Cheap: Construction of Statistical Machine Translation Models with MapReduce
Chris Dyer | Aaron Cordova | Alex Mont | Jimmy Lin
Chris Dyer | Aaron Cordova | Alex Mont | Jimmy Lin
up
Proceedings of the ACL-08: HLT Second Workshop on Syntax and Structure in Statistical Translation (SSST-2)
Proceedings of the ACL-08: HLT Second Workshop on Syntax and Structure in Statistical Translation (SSST-2)
David Chiang | Dekai Wu
David Chiang | Dekai Wu
Imposing Constraints from the Source Tree on ITG Constraints for SMT
Hirofumi Yamamoto | Hideo Okuma | Eiichiro Sumita
Hirofumi Yamamoto | Hideo Okuma | Eiichiro Sumita
A Scalable Decoder for Parsing-Based Machine Translation with Equivalent Language Model State Maintenance
Zhifei Li | Sanjeev Khudanpur
Zhifei Li | Sanjeev Khudanpur
Prior Derivation Models For Formally Syntax-Based Translation Using Linguistically Syntactic Parsing and Tree Kernels
Bowen Zhou | Bing Xiang | Xiaodan Zhu | Yuqing Gao
Bowen Zhou | Bing Xiang | Xiaodan Zhu | Yuqing Gao
Experiments in Discriminating Phrase-Based Translations on the Basis of Syntactic Coupling Features
Vassilina Nikoulina | Marc Dymetman
Vassilina Nikoulina | Marc Dymetman
Multiple Reorderings in Phrase-Based Machine Translation
Niyu Ge | Abe Ittycheriah | Kishore Papineni
Niyu Ge | Abe Ittycheriah | Kishore Papineni
Improving Word Alignment Using Syntactic Dependencies
Yanjun Ma | Sylwia Ozdowska | Yanli Sun | Andy Way
Yanjun Ma | Sylwia Ozdowska | Yanli Sun | Andy Way
up
Proceedings of the Workshop on Current Trends in Biomedical Natural Language Processing
Proceedings of the Workshop on Current Trends in Biomedical Natural Language Processing
Dina Demner-Fushman | Sophia Ananiadou | Kevin Bretonnel Cohen | John Pestian | Jun’ichi Tsujii | Bonnie Webber
Dina Demner-Fushman | Sophia Ananiadou | Kevin Bretonnel Cohen | John Pestian | Jun’ichi Tsujii | Bonnie Webber
A Graph Kernel for Protein-Protein Interaction Extraction
Antti Airola | Sampo Pyysalo | Jari Björne | Tapio Pahikkala | Filip Ginter | Tapio Salakoski
Antti Airola | Sampo Pyysalo | Jari Björne | Tapio Pahikkala | Filip Ginter | Tapio Salakoski
Extracting Clinical Relationships from Patient Narratives
Angus Roberts | Robert Gaizauskas | Mark Hepple
Angus Roberts | Robert Gaizauskas | Mark Hepple
Mining the Biomedical Literature for Genic Information
Catalina O. Tudor | K. Vijay-Shanker | Carl J. Schmidt
Catalina O. Tudor | K. Vijay-Shanker | Carl J. Schmidt
Accelerating the Annotation of Sparse Named Entities by Dynamic Sentence Selection
Yoshimasa Tsuruoka | Jun’ichi Tsujii | Sophia Ananiadou
Yoshimasa Tsuruoka | Jun’ichi Tsujii | Sophia Ananiadou
The BioScope corpus: annotation for negation, uncertainty and their scope in biomedical texts
György Szarvas | Veronika Vincze | Richárd Farkas | János Csirik
György Szarvas | Veronika Vincze | Richárd Farkas | János Csirik
Recognizing Speculative Language in Biomedical Research Articles: A Linguistically Motivated Perspective
Halil Kilicoglu | Sabine Bergler
Halil Kilicoglu | Sabine Bergler
Cascaded Classifiers for Confidence-Based Chemical Named Entity Recognition
Peter Corbett | Ann Copestake
Peter Corbett | Ann Copestake
How to Make the Most of NE Dictionaries in Statistical NER
Yutaka Sasaki | Yoshimasa Tsuruoka | John McNaught | Sophia Ananiadou
Yutaka Sasaki | Yoshimasa Tsuruoka | John McNaught | Sophia Ananiadou
Knowledge Sources for Word Sense Disambiguation of Biomedical Text
Mark Stevenson | Yinkun Guo | Robert Gaizauskas | David Martinez
Mark Stevenson | Yinkun Guo | Robert Gaizauskas | David Martinez
Prediction of Protein Sub-cellular Localization using Information from Texts and Sequences.
Hong-Woo Chun | Chisato Yamasaki | Naomi Saichi | Masayuki Tanaka | Teruyoshi Hishiki | Tadashi Imanishi | Takashi Gojobori | Jin-Dong Kim | Jun’ichi Tsujii | Toshihisa Takagi
Hong-Woo Chun | Chisato Yamasaki | Naomi Saichi | Masayuki Tanaka | Teruyoshi Hishiki | Tadashi Imanishi | Takashi Gojobori | Jin-Dong Kim | Jun’ichi Tsujii | Toshihisa Takagi
A Pilot Annotation to Investigate Discourse Connectivity in Biomedical Text
Hong Yu | Nadya Frid | Susan McRoy | Rashmi Prasad | Alan Lee | Aravind Joshi
Hong Yu | Nadya Frid | Susan McRoy | Rashmi Prasad | Alan Lee | Aravind Joshi
Conditional Random Fields and Support Vector Machines for Disorder Named Entity Recognition in Clinical Texts
Dingcheng Li | Guergana Savova | Karin Kipper-Schuler
Dingcheng Li | Guergana Savova | Karin Kipper-Schuler
Using Natural Language Processing to Classify Suicide Notes
John Pestian | Pawel Matykiewicz | Jacqueline Grupp-Phelan | Sarah Arszman Lavanier | Jennifer Combs | Robert Kowatch
John Pestian | Pawel Matykiewicz | Jacqueline Grupp-Phelan | Sarah Arszman Lavanier | Jennifer Combs | Robert Kowatch
Extracting Protein-Protein Interaction based on Discriminative Training of the Hidden Vector State Model
Deyu Zhou | Yulan He
Deyu Zhou | Yulan He
A preliminary approach to extract drugs by combining UMLS resources and USAN naming conventions
Isabel Segura-Bedmar | Paloma Martínez | Doaa Samy
Isabel Segura-Bedmar | Paloma Martínez | Doaa Samy
CBR-Tagger: a case-based reasoning approach to the gene/protein mention problem
Mariana Neves | Monica Chagoyen | José María Carazo | Alberto Pascual-Montano
Mariana Neves | Monica Chagoyen | José María Carazo | Alberto Pascual-Montano
Determining causal and non-causal relationships in biomedical text by classifying verbs using a Naive Bayesian Classifier
Pieter van der Horn | Bart Bakker | Gijs Geleijnse | Jan Korst | Sergei Kurkin
Pieter van der Horn | Bart Bakker | Gijs Geleijnse | Jan Korst | Sergei Kurkin
Statistical Term Profiling for Query Pattern Mining
Paul Buitelaar | Pinar Oezden Wennerberg | Sonja Zillner
Paul Buitelaar | Pinar Oezden Wennerberg | Sonja Zillner
Using Language Models to Identify Language Impairment in Spanish-English Bilingual Children
Thamar Solorio | Yang Liu
Thamar Solorio | Yang Liu
up
Proceedings of the Tenth Meeting of ACL Special Interest Group on Computational Morphology and Phonology
Proceedings of the Tenth Meeting of ACL Special Interest Group on Computational Morphology and Phonology
Jason Eisner | Jeffrey Heinz
Jason Eisner | Jeffrey Heinz
A Bayesian Model of Natural Language Phonology: Generating Alternations from Underlying Forms
David Ellis
David Ellis
up
Proceedings of the Third Workshop on Innovative Use of NLP for Building Educational Applications
Proceedings of the Third Workshop on Innovative Use of NLP for Building Educational Applications
Joel Tetreault | Jill Burstein | Rachele De Felice
Joel Tetreault | Jill Burstein | Rachele De Felice
Classification Errors in a Domain-Independent Assessment System
Rodney D. Nielsen | Wayne Ward | James H. Martin
Rodney D. Nielsen | Wayne Ward | James H. Martin
Recognizing Noisy Romanized Japanese Words in Learner English
Ryo Nagata | Jun-ichi Kakegawa | Hiromi Sugimoto | Yukiko Yabuta
Ryo Nagata | Jun-ichi Kakegawa | Hiromi Sugimoto | Yukiko Yabuta
An Annotated Corpus Outside Its Original Context: A Corpus-Based Exercise Book
Barbora Hladká | Ondřej Kučera
Barbora Hladká | Ondřej Kučera
Answering Learners’ Questions by Retrieving Question Paraphrases from Social Q&A Sites
Delphine Bernhard | Iryna Gurevych
Delphine Bernhard | Iryna Gurevych
Learner Characteristics and Feedback in Tutorial Dialogue
Kristy Boyer | Robert Phillips | Michael Wallis | Mladen Vouk | James Lester
Kristy Boyer | Robert Phillips | Michael Wallis | Mladen Vouk | James Lester
Automatic Identification of Discourse Moves in Scientific Article Introductions
Nick Pendar | Elena Cotos
Nick Pendar | Elena Cotos
An Analysis of Statistical Models and Features for Reading Difficulty Prediction
Michael Heilman | Kevyn Collins-Thompson | Maxine Eskenazi
Michael Heilman | Kevyn Collins-Thompson | Maxine Eskenazi
Retrieval of Reading Materials for Vocabulary and Reading Practice
Michael Heilman | Le Zhao | Juan Pino | Maxine Eskenazi
Michael Heilman | Le Zhao | Juan Pino | Maxine Eskenazi
Real Time Web Text Classification and Analysis of Reading Difficulty
Eleni Miltsakaki | Audrey Troutt
Eleni Miltsakaki | Audrey Troutt
up
Coling 2008: Proceedings of the workshop Multi-source Multilingual Information Extraction and Summarization
Coling 2008: Proceedings of the workshop Multi-source Multilingual Information Extraction and Summarization
Sivaji Bandyopadhyay | Thierry Poibeau | Horacio Saggion | Roman Yangarber
Sivaji Bandyopadhyay | Thierry Poibeau | Horacio Saggion | Roman Yangarber
Automatic Construction of Domain-specific Dictionaries on Sparse Parallel Corpora in the Nordic languages
Sumithra Velupillai | Hercules Dalianis
Sumithra Velupillai | Hercules Dalianis
Evaluating automatically generated user-focused multi-document summaries for geo-referenced images
Ahmet Aker | Robert Gaizauskas
Ahmet Aker | Robert Gaizauskas
up
Coling 2008: Proceedings of the workshop on Grammar Engineering Across Frameworks
Coling 2008: Proceedings of the workshop on Grammar Engineering Across Frameworks
Stephen Clark | Tracy Holloway King
Stephen Clark | Tracy Holloway King
TuLiPA: Towards a Multi-Formalism Parsing Environment for Grammar Engineering
Laura Kallmeyer | Timm Lichte | Wolfgang Maier | Yannick Parmentier | Johannes Dellert | Kilian Evang
Laura Kallmeyer | Timm Lichte | Wolfgang Maier | Yannick Parmentier | Johannes Dellert | Kilian Evang
Making Speech Look Like Text in the Regulus Development Environment
Elisabeth Kron | Manny Rayner | Marianne Santaholma | Pierrette Bouillon | Agnes Lisowska
Elisabeth Kron | Manny Rayner | Marianne Santaholma | Pierrette Bouillon | Agnes Lisowska
A More Precise Analysis of Punctuation for Broad-Coverage Surface Realization with CCG
Michael White | Rajakrishnan Rajkumar
Michael White | Rajakrishnan Rajkumar
Speeding up LFG Parsing Using C-Structure Pruning
Aoife Cahill | John T. Maxwell III | Paul Meurer | Christian Rohrer | Victoria Rosén
Aoife Cahill | John T. Maxwell III | Paul Meurer | Christian Rohrer | Victoria Rosén
From Grammar-Independent Construction Enumeration to Lexical Types in Computational Grammars
Lars Hellan
Lars Hellan
up
Coling 2008: Proceedings of the Workshop on Cognitive Aspects of the Lexicon (COGALEX 2008)
Coling 2008: Proceedings of the Workshop on Cognitive Aspects of the Lexicon (COGALEX 2008)
Michael Zock | Chu-Ren Huang
Michael Zock | Chu-Ren Huang
Comparing Lexical Relationships Observed within Japanese Collocation Data and Japanese Word Association Norms
Terry Joyce | Irena Srdanović
Terry Joyce | Irena Srdanović
ProPOSEL: a human-oriented prosody and PoS English lexicon for machine-learning and NLP
Claire Brierley | Eric Atwell
Claire Brierley | Eric Atwell
First ideas of user-adapted views of lexicographic data exemplified on OWID and elexiko
Carolin Möller-Spitzer | Christine Möhrs
Carolin Möller-Spitzer | Christine Möhrs
Multilingual Conceptual Access to Lexicon based on Shared Orthography: An ontology-driven study of Chinese and Japanese
Chu-Ren Huang | Ya-Min Chou | Chiyo Hotani | Sheng-Yi Chen | Wan-Ying Lin
Chu-Ren Huang | Ya-Min Chou | Chiyo Hotani | Sheng-Yi Chen | Wan-Ying Lin
Extracting Sense Trees from the Romanian Thesaurus by Sense Segmentation & Dependency Parsing
Neculai Curteanu | Alex Moruz | Diana Trandabăţ
Neculai Curteanu | Alex Moruz | Diana Trandabăţ
Lexical-Functional Correspondences and Their Use in the System of Machine Translation ETAP-3
Andreyeva Sasha
Andreyeva Sasha
The “Close-Distant” Relation of Adjectival Concepts Based on Self-Organizing Map
Kyoko Kanzaki | Noriko Tomuro | Hitoshi Isahara
Kyoko Kanzaki | Noriko Tomuro | Hitoshi Isahara
Toward a cognitive organization for electronic dictionaries, the case for semantic proxemy
Bruno Gaume | Karine Duvignau | Laurent Prévot | Yann Desalle
Bruno Gaume | Karine Duvignau | Laurent Prévot | Yann Desalle
up
Coling 2008: Proceedings of the 3rd Textgraphs workshop on Graph-based Algorithms for Natural Language Processing
Coling 2008: Proceedings of the 3rd Textgraphs workshop on Graph-based Algorithms for Natural Language Processing
Irina Matveeva | Chris Biemann | Monojit Choudhury | Mona Diab
Irina Matveeva | Chris Biemann | Monojit Choudhury | Mona Diab
Acquistion of the Morphological Structure of the Lexicon Based on Lexical Similarity and Formal Analogy
Nabil Hathout
Nabil Hathout
How is Meaning Grounded in Dictionary Definitions?
Alexandre Blondin Massé | Guillaume Chicoisne | Yassine Gargouri | Stevan Harnad | Odile Marcotte | Olivier Picard
Alexandre Blondin Massé | Guillaume Chicoisne | Yassine Gargouri | Stevan Harnad | Odile Marcotte | Olivier Picard
Encoding Tree Pair-Based Graphs in Learning Algorithms: The Textual Entailment Recognition Case
Alessandro Moschitti | Fabio Massimo Zanzotto
Alessandro Moschitti | Fabio Massimo Zanzotto
Graph-Based Clustering for Semantic Classification of Onomatopoetic Words
Kenichi Ichioka | Fumiyo Fukumoto
Kenichi Ichioka | Fumiyo Fukumoto
Semantic Structure from Correspondence Analysis
Barbara McGillivray | Christer Johansson | Daniel Apollon
Barbara McGillivray | Christer Johansson | Daniel Apollon
up
Proceedings of the Ninth International Workshop on Tree Adjoining Grammar and Related Frameworks (TAG+9)
Proceedings of the Ninth International Workshop on Tree Adjoining Grammar and Related Frameworks (TAG+9)
Claire Gardent | Anoop Sarkar
Claire Gardent | Anoop Sarkar
Compositional Semantics of Coordination using Synchronous Tree Adjoining Grammar
Chung-hye Han | David Potter | Dennis R. Storoshenko
Chung-hye Han | David Potter | Dennis R. Storoshenko
Synchronous Vector TAG for Syntax and Semantics: Control Verbs, Relative Clauses, and Inverse Linking
Rebecca Nesson | Stuart Shieber
Rebecca Nesson | Stuart Shieber
Modeling Mobile Intention Recognition Problems with Spatially Constrained Tree-Adjoining Grammars
Peter Kiefer
Peter Kiefer
TuLiPA: A syntax-semantics parsing environment for mildly context-sensitive formalisms
Yannick Parmentier | Laura Kallmeyer | Wolfgang Maier | Timm Lichte | Johannes Dellert
Yannick Parmentier | Laura Kallmeyer | Wolfgang Maier | Timm Lichte | Johannes Dellert
up
Proceedings of the 5th International Workshop on Spoken Language Translation: Evaluation Campaign
This paper gives an overview of the evaluation campaign results of the International1Workshop on Spoken Language Translation (IWSLT) 2008 . In this workshop, we focused on the translation of spontaneous speech recorded in a real situation and the feasability of pivot-language-based translation approaches. The translation directions were English into Chinese and vice versa for the Challenge Task, Chinese into English and English into Spanish for the Pivot Task, and Arabic, Chinese, Spanish into English for the standard BTEC Task. In total, 19 research groups building 58 MT engines participated in this year’s event. Automatic and subjective evaluations were carried out in order to investigate the impact of spontaneity aspects of field data experiments on automatic speech recognition (ASR) and machine translation (MT) system performance as well as the robustness of state-of-the-art MT systems towards speech-to-speech translation in real environments.
The CMU syntax-augmented machine translation system: SAMT on Hadoop with n-best alignments.
Andreas Zollmann | Ashish Venugopal | Stephan Vogel
Andreas Zollmann | Ashish Venugopal | Stephan Vogel
We present the CMU Syntax Augmented Machine Translation System that was used in the IWSLT-08 evaluation campaign. We participated in the Full-BTEC data track for Chinese-English translation, focusing on transcript translation. For this year’s evaluation, we ported the Syntax Augmented MT toolkit [1] to the Hadoop MapReduce [2] parallel processing architecture, allowing us to efficiently run experiments evaluating a novel “wider pipelines” approach to integrate evidence from N -best alignments into our translation models. We describe each step of the MapReduce pipeline as it is implemented in the open-source SAMT toolkit, and show improvements in translation quality by using N-best alignments in both hierarchical and syntax augmented translation systems.
Exploiting alignment techniques in MATREX: the DCU machine translation system for IWSLT 2008.
Yanjun Ma | John Tinsley | Hany Hassan | Jinhua Du | Andy Way
Yanjun Ma | John Tinsley | Hany Hassan | Jinhua Du | Andy Way
In this paper, we give a description of the machine translation (MT) system developed at DCU that was used for our third participation in the evaluation campaign of the International Workshop on Spoken Language Translation (IWSLT 2008). In this participation, we focus on various techniques for word and phrase alignment to improve system quality. Specifically, we try out our word packing and syntax-enhanced word alignment techniques for the Chinese–English task and for the English–Chinese task for the first time. For all translation tasks except Arabic–English, we exploit linguistically motivated bilingual phrase pairs extracted from parallel treebanks. We smooth our translation tables with out-of-domain word translations for the Arabic–English and Chinese–English tasks in order to solve the problem of the high number of out of vocabulary items. We also carried out experiments combining both in-domain and out-of-domain data to improve system performance and, finally, we deploy a majority voting procedure combining a language model-based method and a translation-based method for case and punctuation restoration. We participated in all the translation tasks and translated both the single-best ASR hypotheses and the correct recognition results. The translation results confirm that our new word and phrase alignment techniques are often helpful in improving translation quality, and the data combination method we proposed can significantly improve system performance.
This paper reports on the participation of FBK at the IWSLT 2008 Evaluation. Main effort has been spent on the Chinese-Spanish Pivot task. We implemented four methods to perform pivot translation. The results on the IWSLT 2008 test data show that our original method for generating training data through random sampling outperforms the best methods based on coupling translation systems. FBK also participated in the Chinese-English Challenge task and the Chinese-English and Chinese-Spanish BTEC tasks, employing the standard state-of-the-art MT system Moses Toolkit.
The GREYC machine translation system for the IWSLT 2008 evaluation campaign.
Yves Lepage | Adrien Lardilleux | Julien Gosme | Jean-Luc Manguin
Yves Lepage | Adrien Lardilleux | Julien Gosme | Jean-Luc Manguin
This year’s GREYC machine translation (MT) system presents three major changes relative to the system presented during the previous campaign, while, of course, remaining a pure example-based MT system that exploits proportional analogies. Firstly, the analogy solver has been replaced with a truly non-deterministic one. Secondly, the engine has been re-engineered and a better control has been introduced. Thirdly, the data used for translation were the data provided by the organizers plus alignments obtained using a new alignment method. This year we chose to have the engine run with the word as the processing unit on the contrary to previous years where the processing unit used to be the character. The tracks the system participated in are all classic BTEC tracks (Arabic-English, Chinese-English and Chinese-Spanish) plus the so-called PIVOT task, where the test set had to be translated from Chinese into Spanish by way of English.
I2R multi-pass machine translation system for IWSLT 2008.
Boxing Chen | Deyi Xiong | Min Zhang | Aiti Aw | Haizhou Li
Boxing Chen | Deyi Xiong | Min Zhang | Aiti Aw | Haizhou Li
In this paper, we describe the system and approach used by the Institute for Infocomm Research (I2R) for the IWSLT 2008 spoken language translation evaluation campaign. In the system, we integrate various decoding algorithms into a multi-pass translation framework. The multi-pass approach enables us to utilize various decoding algorithm and to explore much more hypotheses. This paper reports our design philosophy, overall architecture, each individual system and various system combination methods that we have explored. The performance on development and test sets are reported in detail in the paper. The system has shown competitive performance with respect to the BLEU and METEOR measures in Chinese-English Challenge and BTEC tasks.
The ICT system description for IWSLT 2008.
Yang Liu | Zhongjun He | Haitao Mi | Yun Huang | Yang Feng | Wenbin Jiang | Yajuan Lu | Qun Liu
Yang Liu | Zhongjun He | Haitao Mi | Yun Huang | Yang Feng | Wenbin Jiang | Yajuan Lu | Qun Liu
This paper presents a description for the ICT systems involved in the IWSLT 2008 evaluation campaign. This year, we participated in Chinese-English and English-Chinese translation directions. Four statistical machine translation systems were used: one linguistically syntax-based, two formally syntax-based, and one phrase-based. The outputs of the four SMT systems were fed to a sentence-level system combiner, which was expected to produce better translations than single systems. We will report the results of the four single systems and the combiner on both the development and test sets.
The LIG Arabic/English speech translation system at IWSLT08.
L. Besacier | A. Ben-Youssef | H. Blanchon
L. Besacier | A. Ben-Youssef | H. Blanchon
This paper is a description of the system presented by the LIG laboratory to the IWSLT08 speech translation evaluation. The LIG participated, for the second time this year, in the Arabic to English speech translation task. For translation, we used a conventional statistical phrase-based system developed using the moses open source decoder. We describe chronologically the improvements made since last year, starting from the IWSLT 2007 system, following with the improvements made for our 2008 submission. Then, we discuss in section 5 some post-evaluation experiments made very recently, as well as some on-going work on Arabic / English speech to text translation. This year, the systems were ranked according to the (BLEU+METEOR)/2 score of the primary ASR output run submissions. The LIG was ranked 5th/10 based on this rule.
The LIUM Arabic/English statistical machine translation system for IWSLT 2008.
Holger Schwenk | Yannick Estève | Sadaf Abdul Rauf
Holger Schwenk | Yannick Estève | Sadaf Abdul Rauf
This paper describes the system developed by the LIUM laboratory for the 2008 IWSLT evaluation. We only participated in the Arabic/English BTEC task. We developed a statistical phrase-based system using the Moses toolkit and SYSTRAN’s rule-based translation system to perform a morphological decomposition of the Arabic words. A continuous space language model was deployed to improve the modeling of the target language. Both approaches achieved significant improvements in the BLEU score. The system achieves a score of 49.4 on the test set of the 2008 IWSLT evaluation.
This paper describes the MIT-LL/AFRL statistical MT system and the improvements that were developed during the IWSLT 2008 evaluation campaign. As part of these efforts, we experimented with a number of extensions to the standard phrase-based model that improve performance for both text and speech-based translation on Chinese and Arabic translation tasks. We discuss the architecture of the MIT-LL/AFRL MT system, improvements over our 2007 system, and experiments we ran during the IWSLT-2008 evaluation. Specifically, we focus on 1) novel segmentation models for phrase-based MT, 2) improved lattice and confusion network decoding of speech input, 3) improved Arabic morphology for MT preprocessing, and 4) system combination methods for machine translation.
The NICT/ATR speech translation system for IWSLT 2008.
Masao Utiyama | Andrew Finch | Hideo Okuma | Michael Paul | Hailong Cao | Hirofumi Yamamoto | Keiji Yasuda | Eiichiro Sumita
Masao Utiyama | Andrew Finch | Hideo Okuma | Michael Paul | Hailong Cao | Hirofumi Yamamoto | Keiji Yasuda | Eiichiro Sumita
This paper describes the National Institute of Information and Communications Technology/Advanced Telecommunications Research Institute International (NICT/ATR) statistical machine translation (SMT) system used for the IWSLT 2008 evaluation campaign. We participated in the Chinese–English (Challenge Task), English–Chinese (Challenge Task), Chinese–English (BTEC Task), Chinese–Spanish (BTEC Task), and Chinese–English–Spanish (PIVOT Task) translation tasks. In the English–Chinese translation Challenge Task, we focused on exploring various factors for the English–Chinese translation because the research on the translation of English–Chinese is scarce compared to the opposite direction. In the Chinese–English translation Challenge Task, we employed a novel clustering method, where training sentences similar to the development data in terms of the word error rate formed a cluster. In the pivot translation task, we integrated two strategies for pivot translation by linear interpolation.
The CASIA statistical machine translation system for IWSLT 2008
Yanqing He | Jiajun Zhang | Maoxi Li | Licheng Fang | Yufeng Chen | Yu Zhou | Chengqing Zong
Yanqing He | Jiajun Zhang | Maoxi Li | Licheng Fang | Yufeng Chen | Yu Zhou | Chengqing Zong
This paper describes our statistical machine translation system (CASIA) used in the evaluation campaign of the International Workshop on Spoken Language Translation (IWSLT) 2008. In this year’s evaluation, we participated in challenge task for Chinese-English and English-Chinese, BTEC task for Chinese-English. Here, we mainly introduce the overview of our system, the primary modules, the key techniques, and the evaluation results.
NTT statistical machine translation system for IWSLT 2008.
Katsuhito Sudoh | Taro Watanabe | Jun Suzuki | Hajime Tsukada | Hideki Isozaki
Katsuhito Sudoh | Taro Watanabe | Jun Suzuki | Hajime Tsukada | Hideki Isozaki
The NTT Statistical Machine Translation System consists of two primary components: a statistical machine translation decoder and a reranker. The decoder generates k-best translation canditates using a hierarchical phrase-based translation based on synchronous context-free grammar. The decoder employs a linear feature combination among several real-valued scores on translation and language models. The reranker reorders the k-best translation candidates using Ranking SVMs with a large number of sparse features. This paper describes the two components and presents the results for the evaluation campaign of IWSLT 2008.
POSTECH machine translation system for IWSLT 2008 evaluation campaign.
Jonghoon Lee | Gary Geunbae Lee
Jonghoon Lee | Gary Geunbae Lee
In this paper, we describe POSTECH system for IWSLT 2008 evaluation campaign. The system is based on phrase based statistical machine translation. We set up a baseline system using well known freely available software. A preprocessing method and a language modeling method have been applied to the baseline system in order to improve machine translation quality. The preprocessing method is to identify and remove useless tokens in source texts. And the language modeling method models phrase level n-gram. We have participated in the BTEC tasks to see the effects of our methods.
The QMUL system to the IWSLT 2008 evaluation campaign is a phrase-based statistical MT system implemented in C++. The decoder employs a multi-stack architecture, and uses a beam to manage the search space. We participated in both BTEC Arabic → English and Chinese → English tracks, as well as the PIVOT task. In our first submission to IWSLT, we are particularly interested in seeing how our SMT system performs with speech input, having so far only worked with and translated newswire data sets.
The RWTH machine translation system for IWSLT 2008.
David Vilar | Daniel Stein | Yuqi Zhang | Evgeny Matusov | Arne Mauser | Oliver Bender | Saab Mansour | Hermann Ney
David Vilar | Daniel Stein | Yuqi Zhang | Evgeny Matusov | Arne Mauser | Oliver Bender | Saab Mansour | Hermann Ney
RWTH’s system for the 2008 IWSLT evaluation consists of a combination of different phrase-based and hierarchical statistical machine translation systems. We participated in the translation tasks for the Chinese-to-English and Arabic-to-English language pairs. We investigated different preprocessing techniques, reordering methods for the phrase-based system, including reordering of speech lattices, and syntax-based enhancements for the hierarchical systems. We also tried the combination of the Arabic-to-English and Chinese-to-English outputs as an additional submission.
The TALP&I2R SMT systems for IWSLT 2008.
Maxim Khalilov | Marta R. Costa-jussà | Carlos A. Henríquez Q. | José A. R. Fonollosa | Adolfo Hernández H. | José B. Mariño | Rafael E. Banchs | Chen Boxing | Min Zhang | Aiti Aw | Haizhou Li
Maxim Khalilov | Marta R. Costa-jussà | Carlos A. Henríquez Q. | José A. R. Fonollosa | Adolfo Hernández H. | José B. Mariño | Rafael E. Banchs | Chen Boxing | Min Zhang | Aiti Aw | Haizhou Li
This paper gives a description of the statistical machine translation (SMT) systems developed at the TALP Research Center of the UPC (Universitat Polite`cnica de Catalunya) for our participation in the IWSLT’08 evaluation campaign. We present Ngram-based (TALPtuples) and phrase-based (TALPphrases) SMT systems. The paper explains the 2008 systems’ architecture and outlines translation schemes we have used, mainly focusing on the new techniques that are challenged to improve speech-to-speech translation quality. The novelties we have introduced are: improved reordering method, linear combination of translation and reordering models and new technique dealing with punctuation marks insertion for a phrase-based SMT system. This year we focus on the Arabic-English, Chinese-Spanish and pivot Chinese-(English)-Spanish translation tasks.
The TCH machine translation system for IWSLT 2008.
Haifeng Wang | Hua Wu | Xiaoguang Hu | Zhanyi Liu | Jianfeng Li | Dengjun Ren | Zhengyu Niu
Haifeng Wang | Hua Wu | Xiaoguang Hu | Zhanyi Liu | Jianfeng Li | Dengjun Ren | Zhengyu Niu
This paper reports on the first participation of TCH (Toshiba (China) Research and Development Center) at the IWSLT evaluation campaign. We participated in all the 5 translation tasks with Chinese as source language or target language. For Chinese-English and English-Chinese translation, we used hybrid systems that combine rule-based machine translation (RBMT) method and statistical machine translation (SMT) method. For Chinese-Spanish translation, phrase-based SMT models were used. For the pivot task, we combined the translations generated by a pivot based statistical translation model and a statistical transfer translation model (firstly, translating from Chinese to English, and then from English to Spanish). Moreover, for better performance of MT, we improved each module in the MT systems as follows: adapting Chinese word segmentation to spoken language translation, selecting out-of-domain corpus to build language models, using bilingual dictionaries to correct word alignment results, handling NE translation and selecting translations from the outputs of multiple systems. According to the automatic evaluation results on the full test sets, we top in all the 5 tasks.
Statistical machine translation without long parallel sentences for training data.
Jin’ichi Murakami | Masato Tokuhisa | Satoru Ikehara
Jin’ichi Murakami | Masato Tokuhisa | Satoru Ikehara
In this study, we paid attention to the reliability of phrase table. We have been used the phrase table using Och’s method[2]. And this method sometimes generate completely wrong phrase tables. We found that such phrase table caused by long parallel sentences. Therefore, we removed these long parallel sentences from training data. Also, we utilized general tools for statistical machine translation, such as ”Giza++”[3], ”moses”[4], and ”training-phrase-model.perl”[5]. We obtained a BLEU score of 0.4047 (TEXT) and 0.3553(1-BEST) of the Challenge-EC task for our proposed method. On the other hand, we obtained a BLEU score of 0.3975(TEXT) and 0.3482(1-BEST) of the Challenge-EC task for a standard method. This means that our proposed method was effective for the Challenge-EC task. However, it was not effective for the BTECT-CE and Challenge-CE tasks. And our system was not good performance. For example, our system was the 7th place among 8 system for Challenge-EC task.
The TÜBÍTAK-UEKAE statistical machine translation system for IWSLT 2008.
Coşkun Mermer | Hamza Kaya | Ömer Farukhan Güneş | Mehmet Uğur Doğan
Coşkun Mermer | Hamza Kaya | Ömer Farukhan Güneş | Mehmet Uğur Doğan
We present the TÜBİTAK-UEKAE statistical machine translation system that participated in the IWSLT 2008 evaluation campaign. Our system is based on the open-source phrase-based statistical machine translation software Moses. Additionally, phrase-table augmentation is applied to maximize source language coverage; lexical approximation is applied to replace out-of-vocabulary words with known words prior to decoding; and automatic punctuation insertion is improved. We describe the preprocessing and postprocessing steps and our training and decoding procedures. Results are presented on our participation in the classical Arabic-English and Chinese-English tasks as well as the new Chinese-Spanish direct and Chinese-English-Spanish pivot translation tasks.
up
Proceedings of the 5th International Workshop on Spoken Language Translation: Papers
Phrase-based statistical machine translation with pivot languages.
Nicola Bertoldi | Madalina Barbaiani | Marcello Federico | Roldano Cattoni
Nicola Bertoldi | Madalina Barbaiani | Marcello Federico | Roldano Cattoni
Translation with pivot languages has recently gained attention as a means to circumvent the data bottleneck of statistical machine translation (SMT). This paper tries to give a mathematically sound formulation of the various approaches presented in the literature and introduces new methods for training alignment models through pivot languages. We present experimental results on Chinese-Spanish translation via English, on a popular traveling domain task. In contrast to previous literature, we report experimental results by using parallel corpora that are either disjoint or overlapped on the pivot language side. Finally, our original method for generating training data through random sampling shows to perform as well as the best methods based on the coupling of translation systems.
Improving statistical machine translation by paraphrasing the training data.
Francis Bond | Eric Nichols | Darren Scott Appling | Michael Paul
Francis Bond | Eric Nichols | Darren Scott Appling | Michael Paul
Large amounts of training data are essential for training statistical machine translations systems. In this paper we show how training data can be expanded by paraphrasing one side. The new data is made by parsing then generating using a precise HPSG based grammar, which gives sentences with the same meaning, but minor variations in lexical choice and word order. In experiments with Japanese and English, we showed consistent gains on the Tanaka Corpus with less consistent improvement on the IWSLT 2005 evaluation data.
Evaluating productivity gains of hybrid ASR-MT systems for translation dictation.
Alain Désilets | Marta Stojanovic | Jean-François Lapointe | Rick Rose | Aarthi Reddy
Alain Désilets | Marta Stojanovic | Jean-François Lapointe | Rick Rose | Aarthi Reddy
This paper is about Translation Dictation with ASR, that is, the use of Automatic Speech Recognition (ASR) by human translators, in order to dictate translations. We are particularly interested in the productivity gains that this could provide over conventional keyboard input, and ways in which such gains might be increased through a combination of ASR and Statistical Machine Translation (SMT). In this hybrid technology, the source language text is presented to both the human translator and a SMT system. The latter produces N-best translations hypotheses, which are then used to fine tune the ASR language model and vocabulary towards utterances which are probable translations of source text sentences. We conducted an ergonomic experiment with eight professional translators dictating into French, using a top of the line off-the-shelf ASR system (Dragon NatuallySpeaking 8). We found that the ASR system had an average Word Error Rate (WER) of 11.7 percent, and that translation using this system did not provide statistically significant productivity increases over keyboard input, when following the manufacturer recommended procedure for error correction. However, we found indications that, even in its current imperfect state, French ASR might be beneficial to translators who are already used to dictation (either with ASR or a dictaphone), but more focused experiments are needed to confirm this. We also found that dictation using an ASR with WER of 4 percent or less would have resulted in statistically significant (p less than 0.6) productivity gains in the order of 25.1 percent to 44.9 percent Translated Words Per Minute. We also evaluated the extent to which the limited manufacturer provided Domain Adaptation features could be used to positively bias the ASR using SMT hypotheses. We found that the relative gains in WER were much lower than has been reported in the literature for tighter integration of SMT with ASR, pointing the advantages of tight integration approaches and the need for more research in that area.
Rapid development of an English/Farsi speech-to-speech translation system.
C.-L. Kao | S. Saleem | R. Prasad | F. Choi | P. Natarajan | David Stallard | K. Krstovski | M. Kamali
C.-L. Kao | S. Saleem | R. Prasad | F. Choi | P. Natarajan | David Stallard | K. Krstovski | M. Kamali
Significant advances have been achieved in Speech-to-Speech (S2S) translation systems in recent years. However, rapid configuration of S2S systems for low-resource language pairs and domains remains a challenging problem due to lack of human translated bilingual training data. In this paper, we report on an effort to port our existing English/Iraqi S2S system to the English/Farsi language pair in just 90 days, using only a small amount of training data. This effort included developing acoustic models for Farsi, domain-relevant language models for English and Farsi, and translation models for English-to-Farsi and Farsi-to-English. As part of this work, we developed two novel techniques for expanding the training data, including the reuse of data from different language pairs, and directed collection of new data. In an independent evaluation, the resulting system achieved the highest performance of all systems.
Simultaneous German-English lecture translation.
Muntsin Kolss | Matthias Wölfel | Florian Kraft | Jan Niehues | Matthias Paulik | Alex Waibel
Muntsin Kolss | Matthias Wölfel | Florian Kraft | Jan Niehues | Matthias Paulik | Alex Waibel
In an increasingly globalized world, situations in which people of different native tongues have to communicate with each other become more and more frequent. In many such situations, human interpreters are prohibitively expensive or simply not available. Automatic spoken language translation (SLT), as a cost-effective solution to this dilemma, has received increased attention in recent years. For a broad number of applications, including live SLT of lectures and oral presentations, these automatic systems should ideally operate in real time and with low latency. Large and highly specialized vocabularies as well as strong variations in speaking style – ranging from read speech to free presentations suffering from spontaneous events – make simultaneous SLT of lectures a challenging task. This paper presents our progress in building a simultaneous German-English lecture translation system. We emphasize some of the challenges which are particular to this language pair and propose solutions to tackle some of the problems encountered.
Investigations on large-scale lightly-supervised training for statistical machine translation.
Holger Schwenk
Holger Schwenk
Sentence-aligned bilingual texts are a crucial resource to build statistical machine translation (SMT) systems. In this paper we propose to apply lightly-supervised training to produce additional parallel data. The idea is to translate large amounts of monolingual data (up to 275M words) with an SMT system, and to use those as additional training data. Results are reported for the translation from French into English. We consider two setups: first the intial SMT system is only trained with a very limited amount of human-produced translations, and then the case where we have more than 100 million words. In both conditions, lightly-supervised training achieves significant improvements of the BLEU score.
Analysing soft syntax features and heuristics for hierarchical phrase based machine translation.
David Vilar | Daniel Stein | Hermann Ney
David Vilar | Daniel Stein | Hermann Ney
Similar to phrase-based machine translation, hierarchical systems produce a large proportion of phrases, most of which are supposedly junk and useless for the actual translation. For the hierarchical case, however, the amount of extracted rules is an order of magnitude bigger. In this paper, we investigate several soft constraints in the extraction of hierarchical phrases and whether these help as additional scores in the decoding to prune unneeded phrases. We show the methods that help best.
Improvements in dynamic programming beam search for phrase-based statistical machine translation.
Richard Zens | Hermann Ney
Richard Zens | Hermann Ney
Search is a central component of any statistical machine translation system. We describe the search for phrase-based SMT in detail and show its importance for achieving good translation quality. We introduce an explicit distinction between reordering and lexical hypotheses and organize the pruning accordingly. We show that for the large Chinese-English NIST task already a small number of lexical alternatives is sufficient, whereas a large number of reordering hypotheses is required to achieve good translation quality. The resulting system compares favorably with the current stateof-the-art, in particular we perform a comparison with cube pruning as well as with Moses.
up
Proceedings of the 4th Web as Corpus Workshop
We present an experiment evaluating the contribution of a system called GReG for reranking the snippets returned by Google’s search engine in the 10 best links presented to the user, captured by the use of Google’s API. The evaluation aims at establishing whether or not the introduction of deep linguistic information may improve the accuracy of Google or rather it is the opposite case as maintained by the majority of people working in Information Retrieval, using a Bag Of Words approach. We used 900 questions, answers taken from TREC 8, 9 competitions, execute three different types of evaluation: one without any linguistic aid; a second one with tagging, syntactic constituency contribution; another run with what we call Partial Logical Form. Even though GReG is still work in progress, it is possible to draw clearcut conclusions: adding linguistic information to the evaluation process of the best snippet that can answer a question improves enormously the performance. In another experiment we used the actual associated to the Q/A pairs distributed by one of TREC’s participant, got even higher accuracy.
In this paper, we present GLB, yet another open source, free system to create, exploit linguistic corpora gathered from the web. A simple, robust web crawl algorithm, a multi-dimensional information retrieval tool„ a crude parallelization mechanism are proposed, especially for researchers working in resource-limited environments.
In this paper we present a complete solution for automatic cleaning of arbitrary HTML pages with a goal of using web data as a corpus in the area of natural language processing, computational linguistics. We employ a sequence-labeling approach based on Conditional Random Fields (CRF). Every block of text in analyzed web page is assigned a set of features extracted from the textual content, HTML structure of the page. The blocks are automatically labeled either as content segments containing main web page content, which should be preserved, or as noisy segments not suitable for further linguistic processing, which should be eliminated. Our solution is based on the tool introduced at the CLEANEVAL 2007 shared task workshop. In this paper, we present new CRF features, a handy annotation tool„ new evaluation metrics. Evaluation itself is performed on a random sample of web pages automatically downloaded from the Czech web domain.
Segmenting HTML pages using visual, semantic information
Georgios Petasis | Pavlina Fragkou | Aris Theodorakos | Vangelis Karkaletsis | Constantine D. Spyropoulos
Georgios Petasis | Pavlina Fragkou | Aris Theodorakos | Vangelis Karkaletsis | Constantine D. Spyropoulos
The information explosion of the Web aggravates the problem of effective information retrieval. Even though linguistic approaches found in the literature perform linguistic annotation by creating metadata in the form of tokens, lemmas or part of speech tags, however, this process is insufficient. This is due to the fact that these linguistic metadata do not exploit the actual content of the page, leading to the need of performing semantic annotation based on a predefined semantic model. This paper proposes a new learning approach for performing automatic semantic annotation. This is the result of a two step procedure: the first step partitions a web page into blocks based on its visual layout, while the second, performs subsequent partitioning based on the examination of appearance of specific types of entities denoting the semantic category as well as the application of a number of simple heuristics. Preliminary experiments performed on a manually annotated corpus regarding athletics proved to be very promising.
Identifying near duplicate documents is a challenge often faced in the field of information discovery. Unfortunately many algorithms that find near duplicate pairs of plain text documents perform poorly when used on web pages, where metadata, other extraneous information make that process much more difficult. If the content of the page (e.g., the body of a news article) can be extracted from the page, then the accuracy of the duplicate detection algorithms is greatly increased. Using machine learning techniques to identify the content portion of web pages, we achieve duplicate detection accuracy that is nearly identical to plain text, significantly better than simple heuristic approaches to content extraction. We performed these experiments on a small, but fully annotated corpus.
GlossaNet 2: a linguistic search engine for RSS-based corpora
Cédrick Fairon | Kévin Macé | Hubert Naets
Cédrick Fairon | Kévin Macé | Hubert Naets
This paper presents GlossaNet 2, a free online concordance service that enables users to search into dynamic Web corpora. Two steps are involved in using GlossaNet. At first, users define a corpus by selecting RSS feeds in a preselected pool of sources (they can also add their own RSS feeds). These sources will be visited on a regular basis by a crawler in order to generate a dynamic corpus. Secondly, the user can register one or more search queries on his / her dynamic corpus. Search queries will be re-applied on the corpus every time it is updated, new concordances will be recorded for the user (results can be emailed, published for the user in a privative RSS feed, or they can be viewed online). This service integrates two preexisting software: Corporator (Fairon, 2006), a program that creates corpora by downloading, filtering RSS feeds, Unitex (Paumier, 2003), an open source corpus processor that relies on linguistic resources. After a short introduction, we will briefly present the concept of “RSS corpora”, the assets of this approach to corpus development. We will then give an overview of the GlossaNet architecture, present various cases of use.
Collecting Basque specialized corpora from the web: language-specific performance tweaks, improving topic precision
I. Leturia | I. San Vicente | X. Saralegi | M. Lopez de Lacalle
I. Leturia | I. San Vicente | X. Saralegi | M. Lopez de Lacalle
The de facto standard process for collecting corpora from the Internet (with a given list of words, asking APIs of search engines for random combinations of them, downloading the returned pages) does not give very good precision when searching for texts on a certain topic., this precision is much worse when searching for corpora in the Basque language, due to certain properties inherent in the language, in the Basque web. The method proposed in this paper improves topic precision by using a sample mini-corpus as a basis for the process: the words to be used in the queries are automatically extracted from it„ a final topic-filtering step is performed using document-similarity measures with this sample corpus. We also describe the changes made to the usual process to adapt it to the peculiarities of Basque, alongside other adjustments to improve the general performance of the system, quality of the collected corpora.
Introducing, evaluating ukWaC, a very large web-derived corpus of English
Adriano Ferraresi | Eros Zanchetta | Marco Baroni | Silvia Bernardini
Adriano Ferraresi | Eros Zanchetta | Marco Baroni | Silvia Bernardini
In this paper we introduce ukWaC, a large corpus of English constructed by crawling the .uk Internet domain. The corpus contains more than 2 billion tokens, is one of the largest freely available linguistic resources for English. The paper describes the tools, methodology used in the construction of the corpus, provides a qualitative evaluation of its contents, carried out through a vocabulary-based comparison with the BNC. We conclude by giving practical information about availability, format of the corpus.
The web is the largest available corpus, which could be enormously valuable to many natural language processing applications. However it is becoming very difficult to identify relevant information from the web. We present a system for querying dependency tree collocations from the web. We show its usefulness in identifying relevant information by evaluating its accuracy in the task of extracting classes of named entities. The task achieved a general accuracy of 70%.
up
Proceedings of the Australasian Language Technology Association Workshop 2008
Proceedings of the Australasian Language Technology Association Workshop 2008
Nicola Stokes | David Powers
Nicola Stokes | David Powers
Using Multiple Sources of Agreement Information for Sentiment Classification of Political Transcripts
Clint Burfoot
Clint Burfoot
Classification of Verb Particle Constructions with the Google Web1T Corpus
Jonathan K. Kummerfeld | James R. Curran
Jonathan K. Kummerfeld | James R. Curran
Requests and Commitments in Email are More Complex Than You Think: Eight Reasons to be Cautious
Andrew Lampert | Robert Dale | Cécile Paris
Andrew Lampert | Robert Dale | Cécile Paris
Comparing the Value of Latent Semantic Analysis on two English-to-Indonesian lexical mapping tasks
Eliza Margaretha | Ruli Manurung
Eliza Margaretha | Ruli Manurung
Weighted Mutual Exclusion Bootstrapping for Domain Independent Lexicon and Template Acquisition
Tara McIntosh | James R. Curran
Tara McIntosh | James R. Curran
Investigating Features for Classifying Noun Relations
Dominick Ng | David J. Kedziora | Terry T. W. Miu | James R. Curran
Dominick Ng | David J. Kedziora | Terry T. W. Miu | James R. Curran
up
CoNLL 2008: Proceedings of the Twelfth Conference on Computational Natural Language Learning
CoNLL 2008: Proceedings of the Twelfth Conference on Computational Natural Language Learning
Alexander Clark | Kristina Toutanova
Alexander Clark | Kristina Toutanova
TAG, Dynamic Programming, and the Perceptron for Efficient, Feature-Rich Parsing
Xavier Carreras | Michael Collins | Terry Koo
Xavier Carreras | Michael Collins | Terry Koo
Picking them up and Figuring them out: Verb-Particle Constructions, Noise and Idiomaticity
Carlos Ramisch | Aline Villavicencio | Leonardo Moura | Marco Idiart
Carlos Ramisch | Aline Villavicencio | Leonardo Moura | Marco Idiart
Fast Mapping in Word Learning: What Probabilities Tell Us
Afra Alishahi | Afsaneh Fazly | Suzanne Stevenson
Afra Alishahi | Afsaneh Fazly | Suzanne Stevenson
A MDL-based Model of Gender Knowledge Acquisition
Harmony Marchal | Benoît Lemaire | Maryse Bianco | Philippe Dessus
Harmony Marchal | Benoît Lemaire | Maryse Bianco | Philippe Dessus
Baby SRL: Modeling Early Language Acquisition.
Michael Connor | Yael Gertner | Cynthia Fisher | Dan Roth
Michael Connor | Yael Gertner | Cynthia Fisher | Dan Roth
An Incremental Bayesian Model for Learning Syntactic Categories
Christopher Parisien | Afsaneh Fazly | Suzanne Stevenson
Christopher Parisien | Afsaneh Fazly | Suzanne Stevenson
Fully Unsupervised Graph-Based Discovery of General-Specific Noun Relationships from Web Corpora Frequency Counts
Gaël Dias | Raycho Mukelov | Guillaume Cleuziou
Gaël Dias | Raycho Mukelov | Guillaume Cleuziou
Acquiring Knowledge from the Web to be used as Selectors for Noun Sense Disambiguation
Hansen A. Schwartz | Fernando Gomez
Hansen A. Schwartz | Fernando Gomez
Automatic Chinese Catchword Extraction Based on Time Series Analysis
Han Ren | Donghong Ji | Jing Wan | Lei Han
Han Ren | Donghong Ji | Jing Wan | Lei Han
Easy as ABC? Facilitating Pictorial Communication via Semantically Enhanced Layout
Andrew B. Goldberg | Xiaojin Zhu | Charles R. Dyer | Mohamed Eldawy | Lijie Heng
Andrew B. Goldberg | Xiaojin Zhu | Charles R. Dyer | Mohamed Eldawy | Lijie Heng
A Tree-to-String Phrase-based Model for Statistical Machine Translation
Thai Phuong Nguyen | Akira Shimazu | Tu-Bao Ho | Minh Le Nguyen | Vinh Van Nguyen
Thai Phuong Nguyen | Akira Shimazu | Tu-Bao Ho | Minh Le Nguyen | Vinh Van Nguyen
Trainable Speaker-Based Referring Expression Generation
Giuseppe Di Fabbrizio | Amanda Stent | Srinivas Bangalore
Giuseppe Di Fabbrizio | Amanda Stent | Srinivas Bangalore
The CoNLL 2008 Shared Task on Joint Parsing of Syntactic and Semantic Dependencies
Mihai Surdeanu | Richard Johansson | Adam Meyers | Lluís Màrquez | Joakim Nivre
Mihai Surdeanu | Richard Johansson | Adam Meyers | Lluís Màrquez | Joakim Nivre
A Latent Variable Model of Synchronous Parsing for Syntactic and Semantic Dependencies
James Henderson | Paola Merlo | Gabriele Musillo | Ivan Titov
James Henderson | Paola Merlo | Gabriele Musillo | Ivan Titov
Dependency-based Syntactic–Semantic Analysis with PropBank and NomBank
Richard Johansson | Pierre Nugues
Richard Johansson | Pierre Nugues
Hybrid Learning of Dependency Structures from Heterogeneous Linguistic Resources
Yi Zhang | Rui Wang | Hans Uszkoreit
Yi Zhang | Rui Wang | Hans Uszkoreit
Parsing Syntactic and Semantic Dependencies with Two Single-Stage Maximum Entropy Models
Hai Zhao | Chunyu Kit
Hai Zhao | Chunyu Kit
A Combined Memory-Based Semantic Role Labeler of English
Roser Morante | Walter Daelemans | Vincent Van Asch
Roser Morante | Walter Daelemans | Vincent Van Asch
A Puristic Approach for Joint Dependency Parsing and Semantic Role Labeling
Alexander Volokh | Günter Neumann
Alexander Volokh | Günter Neumann
Discriminative Learning of Syntactic and Semantic Dependencies
Lu Li | Shixi Fan | Xuan Wang | Xiaolong Wang
Lu Li | Shixi Fan | Xuan Wang | Xiaolong Wang
Discriminative vs. Generative Approaches in Semantic Role Labeling
Deniz Yuret | Mehmet Ali Yatbaz | Ahmet Engin Ural
Deniz Yuret | Mehmet Ali Yatbaz | Ahmet Engin Ural
A Pipeline Approach for Syntactic and Semantic Dependency Parsing
Yotaro Watanabe | Masakazu Iwatate | Masayuki Asahara | Yuji Matsumoto
Yotaro Watanabe | Masakazu Iwatate | Masayuki Asahara | Yuji Matsumoto
Semantic Dependency Parsing using N-best Semantic Role Sequences and Roleset Information
Joo-Young Lee | Han-Cheol Cho | Hae-Chang Rim
Joo-Young Lee | Han-Cheol Cho | Hae-Chang Rim
A Cascaded Syntactic and Semantic Dependency Parsing System
Wanxiang Che | Zhenghua Li | Yuxuan Hu | Yongqiang Li | Bing Qin | Ting Liu | Sheng Li
Wanxiang Che | Zhenghua Li | Yuxuan Hu | Yongqiang Li | Bing Qin | Ting Liu | Sheng Li
The Integration of Dependency Relation Classification and Semantic Role Labeling Using Bilayer Maximum Entropy Markov Models
Weiwei Sun | Hongzhan Li | Zhifang Sui
Weiwei Sun | Hongzhan Li | Zhifang Sui
Mixing and Blending Syntactic and Semantic Dependencies
Yvonne Samuelsson | Oscar Täckström | Sumithra Velupillai | Johan Eklund | Mark Fishel | Markus Saers
Yvonne Samuelsson | Oscar Täckström | Sumithra Velupillai | Johan Eklund | Mark Fishel | Markus Saers
Dependency Tree-based SRL with Proper Pruning and Extensive Feature Engineering
Hongling Wang | Honglin Wang | Guodong Zhou | Qiaoming Zhu
Hongling Wang | Honglin Wang | Guodong Zhou | Qiaoming Zhu
up
Proceedings of the Fifth International Natural Language Generation Conference
Proceedings of the Fifth International Natural Language Generation Conference
Michael White | Crystal Nakatsu | David McDonald
Michael White | Crystal Nakatsu | David McDonald
Using Spatial Reference Frames to Generate Grounded Textual Summaries of Georeferenced Data
Ross Turner | Somayajulu Sripada | Ehud Reiter | Ian Davy
Ross Turner | Somayajulu Sripada | Ehud Reiter | Ian Davy
Extractive vs. NLG-based Abstractive Summarization of Evaluative Text: The Effect of Corpus Controversiality
Giuseppe Carenini | Jackie C. K. Cheung
Giuseppe Carenini | Jackie C. K. Cheung
Referring Expressions as Formulas of Description Logic
Carlos Areces | Alexander Koller | Kristina Striegnitz
Carlos Areces | Alexander Koller | Kristina Striegnitz
Attribute Selection for Referring Expression Generation: New Algorithms and Evaluation Methods
Albert Gatt | Anja Belz
Albert Gatt | Anja Belz
Using Tactical NLG to Induce Affective States: Empirical Investigations
Ielka van der Sluis | Chris Mellish
Ielka van der Sluis | Chris Mellish
Automated Metrics That Agree With Human Judgements On Generated Output for an Embodied Conversational Agent
Mary Ellen Foster
Mary Ellen Foster
Simple but effective feedback generation to tutor abstract problem solving
Xin Lu | Barbara Di Eugenio | Stellan Ohlsson | Davide Fossati
Xin Lu | Barbara Di Eugenio | Stellan Ohlsson | Davide Fossati
What’s In a Message? Interpreting Geo-referenced Data for the Visually-impaired
Kavita Thomas | Yaji Sripada
Kavita Thomas | Yaji Sripada
The Effect of Dialogue System Output Style Variation on Users’ Evaluation Judgments and Input Style
Ivana Kruijff-Korbayová | Ciprian Gerstenberger | Olga Kukina | Jan Schehl
Ivana Kruijff-Korbayová | Ciprian Gerstenberger | Olga Kukina | Jan Schehl
The Importance of Narrative and Other Lessons from an Evaluation of an NLG System that Summarises Clinical Data
Ehud Reiter | Albert Gatt | François Portet | Marian van der Meulen
Ehud Reiter | Albert Gatt | François Portet | Marian van der Meulen
Degree of Abstraction in Referring Expression Generation and its Relation with the Construction of the Contrast Set
Raquel Hervás | Pablo Gervás
Raquel Hervás | Pablo Gervás
Parser-Based Retraining for Domain Adaptation of Probabilistic Generators
Deirdre Hogan | Jennifer Foster | Joachim Wagner | Josef van Genabith
Deirdre Hogan | Jennifer Foster | Joachim Wagner | Josef van Genabith
Creation of a New Domain and Evaluation of Comparison Generation in a Natural Language Generation System
Matthew Marge | Amy Isard | Johanna Moore
Matthew Marge | Amy Isard | Johanna Moore
Generating Baseball Summaries from Multiple Perspectives by Reordering Content
Alice Oh | Howard Shrobe
Alice Oh | Howard Shrobe
The GREC Challenge 2008: Overview and Evaluation Results
Anja Belz | Eric Kow | Jette Viethen | Albert Gatt
Anja Belz | Eric Kow | Jette Viethen | Albert Gatt
IS-G: The Comparison of Different Learning Techniques for the Selection of the Main Subject References
Bernd Bohnet
Bernd Bohnet
CNTS: Memory-Based Learning of Generating Repeated References
Iris Hendrickx | Walter Daelemans | Kim Luyckx | Roser Morante | Vincent Van Asch
Iris Hendrickx | Walter Daelemans | Kim Luyckx | Roser Morante | Vincent Van Asch
OSU-2: Generating Referring Expressions with a Maximum Entropy Classifier
Emily Jamison | Dennis Mehay
Emily Jamison | Dennis Mehay
The Fingerprint of Human Referring Expressions and their Surface Realization with Graph Transducers
Bernd Bohnet
Bernd Bohnet
Referring Expression Generation Using Speaker-based Attribute Selection and Trainable Realization (ATTR)
Giuseppe Di Fabbrizio | Amanda J. Stent | Srinivas Bangalore
Giuseppe Di Fabbrizio | Amanda J. Stent | Srinivas Bangalore
NIL-UCM: Most-Frequent-Value-First Attribute Selection and Best-Scoring-Choice Realization
Pablo Gervás | Raquel Hervás | Carlos León
Pablo Gervás | Raquel Hervás | Carlos León
USP-EACH Frequency-based Greedy Attribute Selection for Referring Expressions Generation
Diego Jesus de Lucena | Ivandré Paraboni
Diego Jesus de Lucena | Ivandré Paraboni
Referring Expression Generation Challenge 2008 DIT System Descriptions (DIT-FBI, DIT-TVAS, DIT-CBSR, DIT-RBR, DIT-FBI-CBSR, DIT-TVAS-RBR)
John D. Kelleher | Brian Mac Namee
John D. Kelleher | Brian Mac Namee
up
Semantics in Text Processing. STEP 2008 Conference Proceedings
Combining Knowledge-based Methods and Supervised Learning for Effective Italian Word Sense Disambiguation
Pierpaolo Basile | Marco de Gemmis | Pasquale Lops | Giovanni Semeraro
Pierpaolo Basile | Marco de Gemmis | Pasquale Lops | Giovanni Semeraro
Semantic Representations of Syntactically Marked Discourse Status in Crosslinguistic Perspective
Emily M. Bender | David Goss-Grubbs
Emily M. Bender | David Goss-Grubbs
Augmenting WordNet for Deep Understanding of Text
Peter Clark | Christiane Fellbaum | Jerry R. Hobbs | Phil Harrison | William R. Murray | John Thompson
Peter Clark | Christiane Fellbaum | Jerry R. Hobbs | Phil Harrison | William R. Murray | John Thompson
KnowNet: A Proposal for Building Highly Connected and Dense Knowledge Bases from the Web
Montse Cuadros | German Rigau
Montse Cuadros | German Rigau
Combining Word Sense and Usage for Modeling Frame Semantics
Diego De Cao | Danilo Croce | Marco Pennacchiotti | Roberto Basili
Diego De Cao | Danilo Croce | Marco Pennacchiotti | Roberto Basili
Analyzing the Explanation Structure of Procedural Texts: Dealing with Advice and Warnings
Lionel Fontan | Patrick Saint-Dizier
Lionel Fontan | Patrick Saint-Dizier
From Predicting Predominant Senses to Local Context for Word Sense Disambiguation
Rob Koeling | Diana McCarthy
Rob Koeling | Diana McCarthy
Analysis of ASL Motion Capture Data towards Identification of Verb Type
Evguenia Malaia | John Borneman | Ronnie B. Wilbur
Evguenia Malaia | John Borneman | Ronnie B. Wilbur
Resolving Paraphrases to Support Modeling Language Perception in an Intelligent Agent
Sergei Nirenburg | Marjorie McShane | Stephen Beale
Sergei Nirenburg | Marjorie McShane | Stephen Beale
Refining the Meaning of Sense Labels in PDTB: “Concession”
Livio Robaldo | Eleni Miltsakaki | Jerry R. Hobbs
Livio Robaldo | Eleni Miltsakaki | Jerry R. Hobbs
Connective-based Local Coherence Analysis: A Lexicon for Recognizing Causal Relationships
Manfred Stede
Manfred Stede
Open Knowledge Extraction through Compositional Language Processing
Benjamin Van Durme | Lenhart Schubert
Benjamin Van Durme | Lenhart Schubert
LXGram in the Shared Task “Comparing Semantic Representations” of STEP 2008
António Branco | Francisco Costa
António Branco | Francisco Costa
Baseline Evaluation of WSD and Semantic Dependency in OntoSem
Sergei Nirenburg | Stephen Beale | Marjorie McShane
Sergei Nirenburg | Stephen Beale | Marjorie McShane
Textual Entailment as an Evaluation Framework for Metaphor Resolution: A Proposal
Rodrigo Agerri | John Barnden | Mark Lee | Alan Wallington
Rodrigo Agerri | John Barnden | Mark Lee | Alan Wallington
Representing and Visualizing Calendar Expressions in Texts
Delphine Battistelli | Javier Couto | Jean-Luc Minel | Sylviane R. Schwer
Delphine Battistelli | Javier Couto | Jean-Luc Minel | Sylviane R. Schwer
Addressing the Resource Bottleneck to Create Large-Scale Annotated Texts
Jon Chamberlain | Massimo Poesio | Udo Kruschwitz
Jon Chamberlain | Massimo Poesio | Udo Kruschwitz