Wrapping up. Computational Lexical Semantics. Gemma Boleda 1 Stefan Evert 2. ESSLLI. Bordeaux, France, July 2009.

Size: px
Start display at page:

Download "Wrapping up. Computational Lexical Semantics. Gemma Boleda 1 Stefan Evert 2. ESSLLI. Bordeaux, France, July 2009."

Transcription

1 Computational Lexical Semantics Gemma Boleda 1 Stefan Evert 2 1 Universitat Politècnica de Catalunya 2 University of Osnabrück ESSLLI. Bordeaux, France, July / 18

2 Outline / 18

3 Computational Lexical Semantics inter-annotator studies and Machine Learning approaches to semantic tasks 1 (eventually) improve applications 2 tools in developing semantic theories and getting the basic facts right empirical data tests and measures for hypotheses 3 / 18

4 The role of competitions in Computational Lexical Semantics SemEval, SensEval; shared tasks at CoNLL, EMNLP, ESSLLI,... relevant role: common efforts in the community same problems same datasets (usually made publicly available) gradual increase in difficulty comparability, measure of progress some cons often artificial setup and tasks (necessary step?) focus on particular tasks (WSD, SRL) and not others watch out for the next SemEval in / 18

5 Outline / 18

6 Questionnaire Concepts: hyponymy? thematic roles? selectional restrictions? Resources: ever heard about WordNet? ever browsed through WordNet? FrameNet? PropBank? Tasks: Word Sense Disambiguation? Semantic Role Labeling? Methods and tools: ever done any annotation? Machine Learning experiment? Methodology? ever heard about Weka? 6 / 18

7 Messages The most fundamental problems that empirical computational lexical semantics faces nowadays are due to a lack of theoretical understanding. Lexical semantics not as mature as other fields in linguistics (phonetics, syntax) Conceptual challenges what information to encode how to represent meaning how to represent relations between senses... 7 / 18

8 Messages The most fundamental problems that empirical computational lexical semantics faces nowadays are due to a lack of theoretical understanding. Lexical semantics not as mature as other fields in linguistics (phonetics, syntax) Conceptual challenges what information to encode how to represent meaning how to represent relations between senses... 7 / 18

9 Messages Linguists have a lot to say in that respect but we have to learn how to play the game by the rules Statistical analysis Machine Learning methodology Programming tools and resources: R, Weka, Python,... This course: basics to start playing the game Methodology, resources, tools State of the art and main problems 8 / 18

10 Good news it can be done especially with the help of tools R, weka,... you can use them at several levels as black boxes, to help you gain insight into the problem knowing a bit more knowing a lot innovating (implementing your own (open-source) method?) 9 / 18

11 Good news it can be done especially with the help of tools R, weka,... you can use them at several levels as black boxes, to help you gain insight into the problem knowing a bit more knowing a lot innovating (implementing your own (open-source) method?) 9 / 18

12 Outline / 18

13 Not covered in the course Word Similarity / Word Relatedness WordNet-based measures (intuition: distance in graph) distributional approaches (see course by A. Lenci and S. Evert!) Lexical Acquisition e.g. automatic classification of verbs into semantic classes (induction of frames, Levin classes,... ) syntax-semantics interface Word Relations Acquisition: e.g. automatic detection of hyponyms (Hearst 1992) Labeling: our case study 11 / 18

14 Not covered in the course Tools: formalization! how to represent lexical meaning so it can be used computationally? WordNet: simply through labeled links (hyponymy... ) (formal semantics has not focused on lexical semantics) frameworks: Generative Lexicon, Minimal Recursion Semantics, Ontological Semantics... but no common ground that is common to most linguists (as in phonetics, morphology, syntax) 12 / 18

15 Outline / 18

16 Tools and resources WordNet, FrameNet, VerbNet, OntoNotes, etc. Weka (which you already know now) and Weka book very good introduction to Machine Learning (Part I) very good tutorial of Weka (Part II) 14 / 18

17 If you want to learn how to program we recommend that you use Python it has some drawbacks (regular expressions,... ) but it is high-level, object-oriented, and has several facilities for text processing in conjunction with NLTK Natural Language Toolkit facilities to access corpora and resources some drawbacks: heavily oriented towards symbolic approaches (vs. statistical methods) but useful nevertheless NLTK book just came out (O Reilly) 15 / 18

18 If you want to learn how to program we recommend that you use Python it has some drawbacks (regular expressions,... ) but it is high-level, object-oriented, and has several facilities for text processing in conjunction with NLTK Natural Language Toolkit facilities to access corpora and resources some drawbacks: heavily oriented towards symbolic approaches (vs. statistical methods) but useful nevertheless NLTK book just came out (O Reilly) 15 / 18

19 Accessing WordNet in NLTK 16 / 18

20 And if you want to explore your data corpus processing and querying: Corpus WorkBench! graphics, statistical tests, and more: R! (Links to all these resources and tools soon on the course web page.) 17 / 18

21 Computational Lexical Semantics Gemma Boleda 1 Stefan Evert 2 1 Universitat Politècnica de Catalunya 2 University of Osnabrück ESSLLI. Bordeaux, France, July / 18

Identifying Focus, Techniques and Domain of Scientific Papers

Identifying Focus, Techniques and Domain of Scientific Papers Identifying Focus, Techniques and Domain of Scientific Papers Sonal Gupta Department of Computer Science Stanford University Stanford, CA 94305 sonal@cs.stanford.edu Christopher D. Manning Department of

More information

209 THE STRUCTURE AND USE OF ENGLISH.

209 THE STRUCTURE AND USE OF ENGLISH. 209 THE STRUCTURE AND USE OF ENGLISH. (3) A general survey of the history, structure, and use of the English language. Topics investigated include: the history of the English language; elements of the

More information

Comparing methods for automatic acquisition of Topic Signatures

Comparing methods for automatic acquisition of Topic Signatures Comparing methods for automatic acquisition of Topic Signatures Montse Cuadros, Lluis Padro TALP Research Center Universitat Politecnica de Catalunya C/Jordi Girona, Omega S107 08034 Barcelona {cuadros,

More information

Table of Contents. Chapter No. 1 Introduction 1. iii. xiv. xviii. xix. Page No.

Table of Contents. Chapter No. 1 Introduction 1. iii. xiv. xviii. xix. Page No. Table of Contents Title Declaration by the Candidate Certificate of Supervisor Acknowledgement Abstract List of Figures List of Tables List of Abbreviations Chapter Chapter No. 1 Introduction 1 ii iii

More information

Architecture of an Ontology-Based Domain- Specific Natural Language Question Answering System

Architecture of an Ontology-Based Domain- Specific Natural Language Question Answering System Architecture of an Ontology-Based Domain- Specific Natural Language Question Answering System Athira P. M., Sreeja M. and P. C. Reghuraj Department of Computer Science and Engineering, Government Engineering

More information

Ngram Search Engine with Patterns Combining Token, POS, Chunk and NE Information

Ngram Search Engine with Patterns Combining Token, POS, Chunk and NE Information Ngram Search Engine with Patterns Combining Token, POS, Chunk and NE Information Satoshi Sekine Computer Science Department New York University sekine@cs.nyu.edu Kapil Dalwani Computer Science Department

More information

Comparing Ontology-based and Corpusbased Domain Annotations in WordNet.

Comparing Ontology-based and Corpusbased Domain Annotations in WordNet. Comparing Ontology-based and Corpusbased Domain Annotations in WordNet. A paper by: Bernardo Magnini Carlo Strapparava Giovanni Pezzulo Alfio Glozzo Presented by: rabee ali alshemali Motive. Domain information

More information

ANALYSIS OF LEXICO-SYNTACTIC PATTERNS FOR ANTONYM PAIR EXTRACTION FROM A TURKISH CORPUS

ANALYSIS OF LEXICO-SYNTACTIC PATTERNS FOR ANTONYM PAIR EXTRACTION FROM A TURKISH CORPUS ANALYSIS OF LEXICO-SYNTACTIC PATTERNS FOR ANTONYM PAIR EXTRACTION FROM A TURKISH CORPUS Gürkan Şahin 1, Banu Diri 1 and Tuğba Yıldız 2 1 Faculty of Electrical-Electronic, Department of Computer Engineering

More information

What s in a Lexicon. The Lexicon. Lexicon vs. Dictionary. What kind of Information should a Lexicon contain?

What s in a Lexicon. The Lexicon. Lexicon vs. Dictionary. What kind of Information should a Lexicon contain? What s in a Lexicon What kind of Information should a Lexicon contain? The Lexicon Miriam Butt November 2002 Semantic: information about lexical meaning and relations (thematic roles, selectional restrictions,

More information

Learning Translation Rules from Bilingual English Filipino Corpus

Learning Translation Rules from Bilingual English Filipino Corpus Proceedings of PACLIC 19, the 19 th Asia-Pacific Conference on Language, Information and Computation. Learning Translation s from Bilingual English Filipino Corpus Michelle Wendy Tan, Raymond Joseph Ang,

More information

Chapter 8. Final Results on Dutch Senseval-2 Test Data

Chapter 8. Final Results on Dutch Senseval-2 Test Data Chapter 8 Final Results on Dutch Senseval-2 Test Data The general idea of testing is to assess how well a given model works and that can only be done properly on data that has not been seen before. Supervised

More information

Semantic Relatedness Metric for Wikipedia Concepts Based on Link Analysis and its Application to Word Sense Disambiguation

Semantic Relatedness Metric for Wikipedia Concepts Based on Link Analysis and its Application to Word Sense Disambiguation Semantic Relatedness Metric for Wikipedia Concepts Based on Link Analysis and its Application to Word Sense Disambiguation Denis Turdakov, Pavel Velikhov ISP RAS turdakov@ispras.ru, pvelikhov@yahoo.com

More information

Automatic assignment of Wikipedia encyclopedic entries to WordNet synsets

Automatic assignment of Wikipedia encyclopedic entries to WordNet synsets Automatic assignment of Wikipedia encyclopedic entries to WordNet synsets Maria Ruiz-Casado, Enrique Alfonseca and Pablo Castells Computer Science Dep., Universidad Autonoma de Madrid, 28049 Madrid, Spain

More information

Thesis Proposal Verb Semantics for Natural Language Understanding

Thesis Proposal Verb Semantics for Natural Language Understanding Thesis Proposal Verb Semantics for Natural Language Understanding Derry Tanti Wijaya Abstract A verb is the organizational core of a sentence. Understanding the meaning of the verb is therefore key to

More information

A Mapping of CIDOC CRM Events to German Wordnet for Event Detection in Texts

A Mapping of CIDOC CRM Events to German Wordnet for Event Detection in Texts A Mapping of CIDOC CRM Events to German Wordnet for Event Detection in Texts Martin Scholz Friedrich-Alexander-University Erlangen-Nürnberg Digital Humanities Research Group Outline Motivation: information

More information

Processing: current projects and research at the IXA Group

Processing: current projects and research at the IXA Group Natural Language Processing: current projects and research at the IXA Group IXA Research Group on NLP University of the Basque Country Xabier Artola Zubillaga Motivation A language that seeks to survive

More information

MASTER OF PHILOSOPHY IN ENGLISH AND APPLIED LINGUISTICS

MASTER OF PHILOSOPHY IN ENGLISH AND APPLIED LINGUISTICS University of Cambridge: Programme Specifications Every effort has been made to ensure the accuracy of the information in this programme specification. Programme specifications are produced and then reviewed

More information

Semantic analysis of text and speech

Semantic analysis of text and speech Semantic analysis of text and speech SGN-9206 Signal processing graduate seminar II, Fall 2007 Anssi Klapuri Institute of Signal Processing, Tampere University of Technology, Finland Outline What is semantic

More information

Domain Independent Knowledge Base Population From Structured and Unstructured Data Sources

Domain Independent Knowledge Base Population From Structured and Unstructured Data Sources Proceedings of the Twenty-Fourth International Florida Artificial Intelligence Research Society Conference Domain Independent Knowledge Base Population From Structured and Unstructured Data Sources Michelle

More information

DanNet Teaching and Research Perspectives at CST

DanNet Teaching and Research Perspectives at CST DanNet Teaching and Research Perspectives at CST Patrizia Paggio Centre for Language Technology University of Copenhagen paggio@hum.ku.dk Dias 1 Outline Previous and current research: Concept-based search:

More information

Selected Topics in Applied Machine Learning: An integrating view on data analysis and learning algorithms

Selected Topics in Applied Machine Learning: An integrating view on data analysis and learning algorithms Selected Topics in Applied Machine Learning: An integrating view on data analysis and learning algorithms ESSLLI 2015 Barcelona, Spain http://ufal.mff.cuni.cz/esslli2015 Barbora Hladká hladka@ufal.mff.cuni.cz

More information

Proceedings of the. 6th International Conference on Generative Approaches to the Lexicon

Proceedings of the. 6th International Conference on Generative Approaches to the Lexicon GL2013 Proceedings of the 6th International Conference on Generative Approaches to the Lexicon Generative Lexicon and Distributional Semantics Edited by Roser Saurí, Nicoletta Calzolari, Chu-Ren Huang,

More information

An Artificial Intelligence approach to Arabic and Islamic content on the internet

An Artificial Intelligence approach to Arabic and Islamic content on the internet An Artificial Intelligence approach to Arabic and Islamic content on the internet Eric Atwell, Claire Brierley, Kais Dukes, Majdi Sawalha, Abdul-Baquee Sharaf I-AIBS Institute for Artificial intelligence

More information

An Introduction to Data Mining

An Introduction to Data Mining An Introduction to Intel Beijing wei.heng@intel.com January 17, 2014 Outline 1 DW Overview What is Notable Application of Conference, Software and Applications Major Process in 2 Major Tasks in Detail

More information

Multilingual Event Detection using the NewsReader Pipelines

Multilingual Event Detection using the NewsReader Pipelines Multilingual Event Detection using the NewsReader Pipelines Rodrigo Agerri, Itziar Aldabe, Egoitz Laparra, German Rigau Antske Fokkens 3, Paul Huijgen 3, Ruben Izquierdo 3, Marieke van Erp 3, Piek Vossen

More information

Sentiment Analysis of Movie Reviews and Twitter Statuses. Introduction

Sentiment Analysis of Movie Reviews and Twitter Statuses. Introduction Sentiment Analysis of Movie Reviews and Twitter Statuses Introduction Sentiment analysis is the task of identifying whether the opinion expressed in a text is positive or negative in general, or about

More information

AN OPEN KNOWLEDGE BASE FOR ITALIAN LANGUAGE IN A COLLABORATIVE PERSPECTIVE

AN OPEN KNOWLEDGE BASE FOR ITALIAN LANGUAGE IN A COLLABORATIVE PERSPECTIVE AN OPEN KNOWLEDGE BASE FOR ITALIAN LANGUAGE IN A COLLABORATIVE PERSPECTIVE Chiari I, A. Gangemi, E. Jezek, A. Oltramari, G. Vetere, L. Vieu http://www.sensocomune.it/ Sapienza Università di Roma Université

More information

Master of Arts in Linguistics Syllabus

Master of Arts in Linguistics Syllabus Master of Arts in Linguistics Syllabus Applicants shall hold a Bachelor s degree with Honours of this University or another qualification of equivalent standard from this University or from another university

More information

Bridging CAQDAS with text mining: Text analyst s toolbox for Big Data: Science in the Media Project

Bridging CAQDAS with text mining: Text analyst s toolbox for Big Data: Science in the Media Project Bridging CAQDAS with text mining: Text analyst s toolbox for Big Data: Science in the Media Project Ahmet Suerdem Istanbul Bilgi University; LSE Methodology Dept. Science in the media project is funded

More information

Historical Linguistics. Diachronic Analysis. Two Approaches to the Study of Language. Kinds of Language Change. What is Historical Linguistics?

Historical Linguistics. Diachronic Analysis. Two Approaches to the Study of Language. Kinds of Language Change. What is Historical Linguistics? Historical Linguistics Diachronic Analysis What is Historical Linguistics? Historical linguistics is the study of how languages change over time and of their relationships with other languages. All languages

More information

Phase 2 of the D4 Project. Helmut Schmid and Sabine Schulte im Walde

Phase 2 of the D4 Project. Helmut Schmid and Sabine Schulte im Walde Statistical Verb-Clustering Model soft clustering: Verbs may belong to several clusters trained on verb-argument tuples clusters together verbs with similar subcategorization and selectional restriction

More information

An Overview of Applied Linguistics

An Overview of Applied Linguistics An Overview of Applied Linguistics Edited by: Norbert Schmitt Abeer Alharbi What is Linguistics? It is a scientific study of a language It s goal is To describe the varieties of languages and explain the

More information

Artificial Intelligence

Artificial Intelligence Artificial Intelligence ICS461 Fall 2010 1 Lecture #12B More Representations Outline Logics Rules Frames Nancy E. Reed nreed@hawaii.edu 2 Representation Agents deal with knowledge (data) Facts (believe

More information

Sense-Tagging Verbs in English and Chinese. Hoa Trang Dang

Sense-Tagging Verbs in English and Chinese. Hoa Trang Dang Sense-Tagging Verbs in English and Chinese Hoa Trang Dang Department of Computer and Information Sciences University of Pennsylvania htd@linc.cis.upenn.edu October 30, 2003 Outline English sense-tagging

More information

The Prolog Interface to the Unstructured Information Management Architecture

The Prolog Interface to the Unstructured Information Management Architecture The Prolog Interface to the Unstructured Information Management Architecture Paul Fodor 1, Adam Lally 2, David Ferrucci 2 1 Stony Brook University, Stony Brook, NY 11794, USA, pfodor@cs.sunysb.edu 2 IBM

More information

Module Catalogue for the Bachelor Program in Computational Linguistics at the University of Heidelberg

Module Catalogue for the Bachelor Program in Computational Linguistics at the University of Heidelberg Module Catalogue for the Bachelor Program in Computational Linguistics at the University of Heidelberg March 1, 2007 The catalogue is organized into sections of (1) obligatory modules ( Basismodule ) that

More information

Natural Language Processing. Part 4: lexical semantics

Natural Language Processing. Part 4: lexical semantics Natural Language Processing Part 4: lexical semantics 2 Lexical semantics A lexicon generally has a highly structured form It stores the meanings and uses of each word It encodes the relations between

More information

Testing Data-Driven Learning Algorithms for PoS Tagging of Icelandic

Testing Data-Driven Learning Algorithms for PoS Tagging of Icelandic Testing Data-Driven Learning Algorithms for PoS Tagging of Icelandic by Sigrún Helgadóttir Abstract This paper gives the results of an experiment concerned with training three different taggers on tagged

More information

Building the Multilingual Web of Data: A Hands-on tutorial (ISWC 2014, Riva del Garda - Italy)

Building the Multilingual Web of Data: A Hands-on tutorial (ISWC 2014, Riva del Garda - Italy) Building the Multilingual Web of Data: A Hands-on tutorial (ISWC 2014, Riva del Garda - Italy) Multilingual Word Sense Disambiguation and Entity Linking on the Web based on BabelNet Roberto Navigli, Tiziano

More information

A typology of ontology-based semantic measures

A typology of ontology-based semantic measures A typology of ontology-based semantic measures Emmanuel Blanchard, Mounira Harzallah, Henri Briand, and Pascale Kuntz Laboratoire d Informatique de Nantes Atlantique Site École polytechnique de l université

More information

Discourse Markers in English Writing

Discourse Markers in English Writing Discourse Markers in English Writing Li FENG Abstract Many devices, such as reference, substitution, ellipsis, and discourse marker, contribute to a discourse s cohesion and coherence. This paper focuses

More information

Construction of Thai WordNet Lexical Database from Machine Readable Dictionaries

Construction of Thai WordNet Lexical Database from Machine Readable Dictionaries Construction of Thai WordNet Lexical Database from Machine Readable Dictionaries Patanakul Sathapornrungkij Department of Computer Science Faculty of Science, Mahidol University Rama6 Road, Ratchathewi

More information

Veronika VINCZE, PhD. PERSONAL DATA Date of birth: 1 July 1981 Nationality: Hungarian

Veronika VINCZE, PhD. PERSONAL DATA Date of birth: 1 July 1981 Nationality: Hungarian Veronika VINCZE, PhD CONTACT INFORMATION Hungarian Academy of Sciences Research Group on Artificial Intelligence Tisza Lajos krt. 103., 6720 Szeged, Hungary Phone: +36 62 54 41 40 Mobile: +36 70 22 99

More information

The course is included in the CPD programme for teachers II.

The course is included in the CPD programme for teachers II. Faculties of Humanities and Theology LLYU72, Swedish as a Second Language for Upper Secondary School Teachers, 60 credits Svenska som andraspråk för lärare i gymnasieskolan, 60 högskolepoäng First Cycle

More information

Survey Results: Requirements and Use Cases for Linguistic Linked Data

Survey Results: Requirements and Use Cases for Linguistic Linked Data Survey Results: Requirements and Use Cases for Linguistic Linked Data 1 Introduction This survey was conducted by the FP7 Project LIDER (http://www.lider-project.eu/) as input into the W3C Community Group

More information

Teaching Applied Natural Language Processing: Triumphs and Tribulations

Teaching Applied Natural Language Processing: Triumphs and Tribulations Teaching Applied Natural Language Processing: Triumphs and Tribulations Marti Hearst School of Information Management & Systems University of California, Berkeley Berkeley, CA 94720 hearst@sims.berkeley.edu

More information

D2.4: Two trained semantic decoders for the Appointment Scheduling task

D2.4: Two trained semantic decoders for the Appointment Scheduling task D2.4: Two trained semantic decoders for the Appointment Scheduling task James Henderson, François Mairesse, Lonneke van der Plas, Paola Merlo Distribution: Public CLASSiC Computational Learning in Adaptive

More information

Annotated Corpora in the Cloud: Free Storage and Free Delivery

Annotated Corpora in the Cloud: Free Storage and Free Delivery Annotated Corpora in the Cloud: Free Storage and Free Delivery Graham Wilcock University of Helsinki graham.wilcock@helsinki.fi Abstract The paper describes a technical strategy for implementing natural

More information

Semantic Analysis of. Tag Similarity Measures in. Collaborative Tagging Systems

Semantic Analysis of. Tag Similarity Measures in. Collaborative Tagging Systems Semantic Analysis of Tag Similarity Measures in Collaborative Tagging Systems 1 Ciro Cattuto, 2 Dominik Benz, 2 Andreas Hotho, 2 Gerd Stumme 1 Complex Networks Lagrange Laboratory (CNLL), ISI Foundation,

More information

Czech Verbs of Communication and the Extraction of their Frames

Czech Verbs of Communication and the Extraction of their Frames Czech Verbs of Communication and the Extraction of their Frames Václava Benešová and Ondřej Bojar Institute of Formal and Applied Linguistics ÚFAL MFF UK, Malostranské náměstí 25, 11800 Praha, Czech Republic

More information

Chapter 2 Senso Comune: A Collaborative Knowledge Resource for Italian

Chapter 2 Senso Comune: A Collaborative Knowledge Resource for Italian Chapter 2 Senso Comune: A Collaborative Knowledge Resource for Italian Alessandro Oltramari, Guido Vetere, Isabella Chiari, Elisabetta Jezek, Fabio Massimo Zanzotto, Malvina Nissim, and Aldo Gangemi Abstract

More information

Text Mining: The state of the art and the challenges

Text Mining: The state of the art and the challenges Text Mining: The state of the art and the challenges Ah-Hwee Tan Kent Ridge Digital Labs 21 Heng Mui Keng Terrace Singapore 119613 Email: ahhwee@krdl.org.sg Abstract Text mining, also known as text data

More information

Title: Chinese Characters and Top Ontology in EuroWordNet

Title: Chinese Characters and Top Ontology in EuroWordNet Title: Chinese Characters and Top Ontology in EuroWordNet Paper by: Shun Sylvia Wong & Karel Pala Presentation By: Patrick Baker Introduction WordNet, Cyc, HowNet, and EuroWordNet each use a hierarchical

More information

Empirical Machine Translation and its Evaluation

Empirical Machine Translation and its Evaluation Empirical Machine Translation and its Evaluation EAMT Best Thesis Award 2008 Jesús Giménez (Advisor, Lluís Màrquez) Universitat Politècnica de Catalunya May 28, 2010 Empirical Machine Translation Empirical

More information

TRINITY COLLEGE. An Investigation of Sentiment Analysis for Political News Feeds

TRINITY COLLEGE. An Investigation of Sentiment Analysis for Political News Feeds University of Dublin TRINITY COLLEGE An Investigation of Sentiment Analysis for Political News Feeds Seán O Sullivan B.A. (Mod.) Computer Science Final Year Project April 2011 Supervisor: Prof. Khurshid

More information

Doctoral Consortium 2013 Dept. Lenguajes y Sistemas Informáticos UNED

Doctoral Consortium 2013 Dept. Lenguajes y Sistemas Informáticos UNED Doctoral Consortium 2013 Dept. Lenguajes y Sistemas Informáticos UNED 17 19 June 2013 Monday 17 June Salón de Actos, Facultad de Psicología, UNED 15.00-16.30: Invited talk Eneko Agirre (Euskal Herriko

More information

VCU-TSA at Semeval-2016 Task 4: Sentiment Analysis in Twitter

VCU-TSA at Semeval-2016 Task 4: Sentiment Analysis in Twitter VCU-TSA at Semeval-2016 Task 4: Sentiment Analysis in Twitter Gerard Briones and Kasun Amarasinghe and Bridget T. McInnes, PhD. Department of Computer Science Virginia Commonwealth University Richmond,

More information

PARMA: A Predicate Argument Aligner

PARMA: A Predicate Argument Aligner PARMA: A Predicate Argument Aligner Travis Wolfe, Benjamin Van Durme, Mark Dredze, Nicholas Andrews, Charley Beller, Chris Callison-Burch, Jay DeYoung, Justin Snyder, Jonathan Weese, Tan Xu, and Xuchen

More information

Visualizing WordNet Structure

Visualizing WordNet Structure Visualizing WordNet Structure Jaap Kamps Abstract Representations in WordNet are not on the level of individual words or word forms, but on the level of word meanings (lexemes). A word meaning, in turn,

More information

Motivation. Korpus-Abfrage: Werkzeuge und Sprachen. Overview. Languages of Corpus Query. SARA Query Possibilities 1

Motivation. Korpus-Abfrage: Werkzeuge und Sprachen. Overview. Languages of Corpus Query. SARA Query Possibilities 1 Korpus-Abfrage: Werkzeuge und Sprachen Gastreferat zur Vorlesung Korpuslinguistik mit und für Computerlinguistik Charlotte Merz 3. Dezember 2002 Motivation Lizentiatsarbeit: A Corpus Query Tool for Automatically

More information

Introduction to formal semantics -

Introduction to formal semantics - Introduction to formal semantics - Introduction to formal semantics 1 / 25 structure Motivation - Philosophy paradox antinomy division in object und Meta language Semiotics syntax semantics Pragmatics

More information

How To Understand The History Of The Semantic Web

How To Understand The History Of The Semantic Web Semantic Web in 10 years: Semantics with a purpose [Position paper for the ISWC2012 Workshop on What will the Semantic Web look like 10 years from now? ] Marko Grobelnik, Dunja Mladenić, Blaž Fortuna J.

More information

Dealing with digital Information richness in supply chain Management - A review and a Big Data Analytics approach

Dealing with digital Information richness in supply chain Management - A review and a Big Data Analytics approach Florian Kache Dealing with digital Information richness in supply chain Management - A review and a Big Data Analytics approach kassel IH university press Contents Acknowledgements Preface Glossary Figures

More information

Customer Intentions Analysis of Twitter Based on Semantic Patterns

Customer Intentions Analysis of Twitter Based on Semantic Patterns Customer Intentions Analysis of Twitter Based on Semantic Patterns Mohamed Hamroun mohamed.hamrounn@gmail.com Mohamed Salah Gouider ms.gouider@yahoo.fr Lamjed Ben Said lamjed.bensaid@isg.rnu.tn ABSTRACT

More information

EDUCATIONAL REGULATION OF THE MASTER S DEGREE COURSE IN COGNITIVE SCIENCE

EDUCATIONAL REGULATION OF THE MASTER S DEGREE COURSE IN COGNITIVE SCIENCE EDUCATIONAL REGULATION OF THE MASTER S DEGREE COURSE IN COGNITIVE SCIENCE CONTENTS Title I - Establishment and start-up... 3 Art. 1 General information... 3 Art. 2 - Initiatives for quality assurance...

More information

Domain Knowledge Extracting in a Chinese Natural Language Interface to Databases: NChiql

Domain Knowledge Extracting in a Chinese Natural Language Interface to Databases: NChiql Domain Knowledge Extracting in a Chinese Natural Language Interface to Databases: NChiql Xiaofeng Meng 1,2, Yong Zhou 1, and Shan Wang 1 1 College of Information, Renmin University of China, Beijing 100872

More information

ONTOLOGIES A short tutorial with references to YAGO Cosmina CROITORU

ONTOLOGIES A short tutorial with references to YAGO Cosmina CROITORU ONTOLOGIES p. 1/40 ONTOLOGIES A short tutorial with references to YAGO Cosmina CROITORU Unlocking the Secrets of the Past: Text Mining for Historical Documents Blockseminar, 21.2.-11.3.2011 ONTOLOGIES

More information

LabelTranslator - A Tool to Automatically Localize an Ontology

LabelTranslator - A Tool to Automatically Localize an Ontology LabelTranslator - A Tool to Automatically Localize an Ontology Mauricio Espinoza 1, Asunción Gómez Pérez 1, and Eduardo Mena 2 1 UPM, Laboratorio de Inteligencia Artificial, 28660 Boadilla del Monte, Spain

More information

The use of binary codes to represent characters

The use of binary codes to represent characters The use of binary codes to represent characters Teacher s Notes Lesson Plan x Length 60 mins Specification Link 2.1.4/hi Character Learning objective (a) Explain the use of binary codes to represent characters

More information

The Knowledge Sharing Infrastructure KSI. Steven Krauwer

The Knowledge Sharing Infrastructure KSI. Steven Krauwer The Knowledge Sharing Infrastructure KSI Steven Krauwer 1 Why a KSI? Building or using a complex installation requires specialized skills and expertise. CLARIN is no exception. CLARIN is populated with

More information

Self-Monitoring in Social Networks

Self-Monitoring in Social Networks Self-Monitoring in Social Networks Amin Anjomshoaa 1, Khue Vo Sao 1, Amirreza Tahamtan 1, A Min Tjoa 1, Edgar Weippl 2 1 Institute of Software Technology and Interactive Systems, Vienna University of Technology

More information

Teaching Formal Methods for Computational Linguistics at Uppsala University

Teaching Formal Methods for Computational Linguistics at Uppsala University Teaching Formal Methods for Computational Linguistics at Uppsala University Roussanka Loukanova Computational Linguistics Dept. of Linguistics and Philology, Uppsala University P.O. Box 635, 751 26 Uppsala,

More information

Using Use Cases for requirements capture. Pete McBreen. 1998 McBreen.Consulting

Using Use Cases for requirements capture. Pete McBreen. 1998 McBreen.Consulting Using Use Cases for requirements capture Pete McBreen 1998 McBreen.Consulting petemcbreen@acm.org All rights reserved. You have permission to copy and distribute the document as long as you make no changes

More information

Learning is a very general term denoting the way in which agents:

Learning is a very general term denoting the way in which agents: What is learning? Learning is a very general term denoting the way in which agents: Acquire and organize knowledge (by building, modifying and organizing internal representations of some external reality);

More information

Does it fit? KOS evaluation using the ICE-Map Visualization.

Does it fit? KOS evaluation using the ICE-Map Visualization. Does it fit? KOS evaluation using the ICE-Map Visualization. Kai Eckert 1, Dominique Ritze 1, and Magnus Pfeffer 2 1 University of Mannheim University Library Mannheim, Germany {kai.eckert,dominique.ritze}@bib.uni-mannheim.de

More information

Semantic annotation of requirements for automatic UML class diagram generation

Semantic annotation of requirements for automatic UML class diagram generation www.ijcsi.org 259 Semantic annotation of requirements for automatic UML class diagram generation Soumaya Amdouni 1, Wahiba Ben Abdessalem Karaa 2 and Sondes Bouabid 3 1 University of tunis High Institute

More information

Session 15 OF, Unpacking the Actuary's Technical Toolkit. Moderator: Albert Jeffrey Moore, ASA, MAAA

Session 15 OF, Unpacking the Actuary's Technical Toolkit. Moderator: Albert Jeffrey Moore, ASA, MAAA Session 15 OF, Unpacking the Actuary's Technical Toolkit Moderator: Albert Jeffrey Moore, ASA, MAAA Presenters: Melissa Boudreau, FCAS Albert Jeffrey Moore, ASA, MAAA Christopher Kenneth Peek Yonasan Schwartz,

More information

Learning Domain Ontologies from Document Warehouses and Dedicated Web Sites

Learning Domain Ontologies from Document Warehouses and Dedicated Web Sites Learning Domain Ontologies from Document Warehouses and Dedicated Web Sites Roberto Navigli Università di Roma La Sapienza Paola Velardi Università di Roma La Sapienza We present a method and a tool, OntoLearn,

More information

Picking them up and Figuring them out: Verb-Particle Constructions, Noise and Idiomaticity

Picking them up and Figuring them out: Verb-Particle Constructions, Noise and Idiomaticity Picking them up and Figuring them out: Verb-Particle Constructions, Noise and Idiomaticity Carlos Ramisch, Aline Villavicencio, Leonardo Moura and Marco Idiart Institute of Informatics, Federal University

More information

An Ontological Document Management System

An Ontological Document Management System An Ontological Document Management System Eric Simon, Iulian Ciorăscu, and Kilian Stoffel Information Management Institute, University of Neuchâtel, Switzerland, {eric.simon iulian.ciorascu kilian.stoffel}@unine.ch,

More information

Knowledge-Based WSD on Specific Domains: Performing Better than Generic Supervised WSD

Knowledge-Based WSD on Specific Domains: Performing Better than Generic Supervised WSD Knowledge-Based WSD on Specific Domains: Performing Better than Generic Supervised WSD Eneko Agirre and Oier Lopez de Lacalle and Aitor Soroa Informatika Fakultatea, University of the Basque Country 20018,

More information

Master of Arts in Teaching English to Speakers of Other Languages (MA TESOL)

Master of Arts in Teaching English to Speakers of Other Languages (MA TESOL) Master of Arts in Teaching English to Speakers of Other Languages (MA TESOL) Overview Teaching English to non-native English speakers requires skills beyond just knowing the language. Teachers must have

More information

SEMANTIC VIDEO ANNOTATION IN E-LEARNING FRAMEWORK

SEMANTIC VIDEO ANNOTATION IN E-LEARNING FRAMEWORK SEMANTIC VIDEO ANNOTATION IN E-LEARNING FRAMEWORK Antonella Carbonaro, Rodolfo Ferrini Department of Computer Science University of Bologna Mura Anteo Zamboni 7, I-40127 Bologna, Italy Tel.: +39 0547 338830

More information

Quality Translation:!

Quality Translation:! LAUNCH PAD LaunchPad Workshop 2013 LAUNCHPAD tand 02.07.2012 Quality Translation:! Addressing the Next Barrier! to Multilingual Communication! on the Internet! Hans Uszkoreit DFKI!! A Concern: Language

More information

Course Syllabus For Operations Management. Management Information Systems

Course Syllabus For Operations Management. Management Information Systems For Operations Management and Management Information Systems Department School Year First Year First Year First Year Second year Second year Second year Third year Third year Third year Third year Third

More information

Isabelle Debourges, Sylvie Guilloré-Billot, Christel Vrain

Isabelle Debourges, Sylvie Guilloré-Billot, Christel Vrain /HDUQLQJ9HUEDO5HODWLRQVLQ7H[W0DSV Isabelle Debourges, Sylvie Guilloré-Billot, Christel Vrain LIFO Rue Léonard de Vinci 45067 Orléans cedex 2 France email: {debourge, billot, christel.vrain}@lifo.univ-orleans.fr

More information

SemEval-2015 Task 15: A Corpus Pattern Analysis Dictionary-Entry-Building Task

SemEval-2015 Task 15: A Corpus Pattern Analysis Dictionary-Entry-Building Task SemEval-2015 Task 15: A Corpus Pattern Analysis Dictionary-Entry-Building Task Vít Baisa Masaryk University xbaisa@fi.muni.cz Ismaïl El Maarouf University of Wolverhampton i.el-maarouf@wlv.ac.uk Jane Bradbury

More information

Converging Web-Data and Database Data: Big - and Small Data via Linked Data

Converging Web-Data and Database Data: Big - and Small Data via Linked Data DBKDA/WEB Panel 2014, Chamonix, 24.04.2014 DBKDA/WEB Panel 2014, Chamonix, 24.04.2014 Reutlingen University Converging Web-Data and Database Data: Big - and Small Data via Linked Data Moderation: Fritz

More information

How the Computer Translates. Svetlana Sokolova President and CEO of PROMT, PhD.

How the Computer Translates. Svetlana Sokolova President and CEO of PROMT, PhD. Svetlana Sokolova President and CEO of PROMT, PhD. How the Computer Translates Machine translation is a special field of computer application where almost everyone believes that he/she is a specialist.

More information

Overview of MT techniques. Malek Boualem (FT)

Overview of MT techniques. Malek Boualem (FT) Overview of MT techniques Malek Boualem (FT) This section presents an standard overview of general aspects related to machine translation with a description of different techniques: bilingual, transfer,

More information

Psychology G4470. Psychology and Neuropsychology of Language. Spring 2013.

Psychology G4470. Psychology and Neuropsychology of Language. Spring 2013. Psychology G4470. Psychology and Neuropsychology of Language. Spring 2013. I. Course description, as it will appear in the bulletins. II. A full description of the content of the course III. Rationale

More information

Exploiting Comparable Corpora and Bilingual Dictionaries. the Cross Language Text Categorization

Exploiting Comparable Corpora and Bilingual Dictionaries. the Cross Language Text Categorization Exploiting Comparable Corpora and Bilingual Dictionaries for Cross-Language Text Categorization Alfio Gliozzo and Carlo Strapparava ITC-Irst via Sommarive, I-38050, Trento, ITALY {gliozzo,strappa}@itc.it

More information

Developing a large semantically annotated corpus

Developing a large semantically annotated corpus Developing a large semantically annotated corpus Valerio Basile, Johan Bos, Kilian Evang, Noortje Venhuizen Center for Language and Cognition Groningen (CLCG) University of Groningen The Netherlands {v.basile,

More information

TERMINOGRAPHY and LEXICOGRAPHY What is the difference? Summary. Anja Drame TermNet

TERMINOGRAPHY and LEXICOGRAPHY What is the difference? Summary. Anja Drame TermNet TERMINOGRAPHY and LEXICOGRAPHY What is the difference? Summary Anja Drame TermNet Summary/ Conclusion Variety of language (GPL = general purpose SPL = special purpose) Lexicography GPL SPL (special-purpose

More information

Journée Thématique Big Data 13/03/2015

Journée Thématique Big Data 13/03/2015 Journée Thématique Big Data 13/03/2015 1 Agenda About Flaminem What Do We Want To Predict? What Is The Machine Learning Theory Behind It? How Does It Work In Practice? What Is Happening When Data Gets

More information

An empirical study of semantic similarity in WordNet and Word2Vec. A Thesis

An empirical study of semantic similarity in WordNet and Word2Vec. A Thesis An empirical study of semantic similarity in WordNet and Word2Vec A Thesis Submitted to the Graduate Faculty of the University of New Orleans in partial fulfillment of the requirements for the degree of

More information

Masters in Human Computer Interaction

Masters in Human Computer Interaction Masters in Human Computer Interaction Programme Requirements Taught Element, and PG Diploma in Human Computer Interaction: 120 credits: IS5101 CS5001 CS5040 CS5041 CS5042 or CS5044 up to 30 credits from

More information

PyCantonese: Cantonese linguistic research in the age of big data

PyCantonese: Cantonese linguistic research in the age of big data PyCantonese: Cantonese linguistic research in the age of big data Jackson L. Lee University of Chicago http://jacksonllee.com Childhood Bilingualism Research Center, CUHK September 15, 2015 Grammar versus

More information

Experiments in Web Page Classification for Semantic Web

Experiments in Web Page Classification for Semantic Web Experiments in Web Page Classification for Semantic Web Asad Satti, Nick Cercone, Vlado Kešelj Faculty of Computer Science, Dalhousie University E-mail: {rashid,nick,vlado}@cs.dal.ca Abstract We address

More information

Qualitative Research. A primer. Developed by: Vicki L. Wise, Ph.D. Portland State University

Qualitative Research. A primer. Developed by: Vicki L. Wise, Ph.D. Portland State University Qualitative Research A primer Developed by: Vicki L. Wise, Ph.D. Portland State University Overview In this session, we will investigate qualitative research methods. At the end, I am hopeful that you

More information