A Relation Extraction Method between Related Concepts using Web Search

Size: px
Start display at page:

Download "A Relation Extraction Method between Related Concepts using Web Search"

Transcription

1 DEIM Forum 2010 C1-2 Web Web Wikipedia is-a a-part-of Wikipedia Web Wikipedia Web Wikipedia Abstract A Relation Extraction Method between Related Concepts using Web Search Masumi SHIRAKAWA, Kotaro NAKAYAMA, Eiji ARAMAKI, Takahiro HARA, and Shojiro NISHIO Department of Multimedia Engineering, Graduate School of Information Science and Technology, Osaka University 1-5 Yamadaoka, Suita, Osaka , Japan The Center for Knowledge Structuring, The University of Tokyo Hongo, Bunkyo-ku, Tokyo , Japan {shirakawa.masumi,hara,nishio}@ist.osaka-u.ac.jp, nakayama@cks.u-tokyo.ac.jp, eiji.aramaki@gmail.com Construction of a huge scale ontology covering many named entities, domain specific terms and relations among these concepts is one of the essential technologies in the next generation Web based on semantics. In this work, we aim at automated ontology construction with a wide coverage of concepts and these relations by combining information on the Web with Wikipedia. In this paper, we propose a relation extraction method which gets pairs of co-related concepts from an association thesaurus extracted from Wikipedia and extracts relations by using Web search. Key words relation extraction, natural language processing (NLP), thesaurus, Wikipedia 1

2 1. Web [1] [2] OpenCyc 1 WordNet [3] EDR 2 is-a a-part-of Web Wikipedia Wikipedia Wikipedia Web Wikipedia [4], [5] Web Web Web 2. Wikipedia Web Wikipedia Wiki Web Web Fig. 1 Ontology reconstruction from Case Frames Wikipedia URL [6] Wikipedia DBpedia [7] YAGO [8], [9] DBpedia Wikipedia RDF (Resource Description Framework) YAGO WordNet Wikipedia WordNet [10] [11] Wikipedia is-a SVM (Support Vector Machine) Wikipedia Wikipedia Web is-a a-part-of [12], [13] 1 2

3 1 R Table 1 A list of link texts of Softbank and reliability R 2 Wikipedia Fig. 2 An example of Wikipedia Thesaurus [14] Web [15] Web 3. Web 3. 1 Web Wikipedia [4], [5] Wikipedia Wikipedia 2 iphone ipod touch iphone BlackBerry Wikipedia Web Wikipedia SoftBank SOFTBANK Softbank Web Web Web Web Wikipedia [16] Wikipedia Wikipedia 1 1 c L c l i L c M c (l i ) l i R c (l i ) 3

4 3 Web Fig. 3 An example of query generation with case particles Fig. 4 4 An example of relation extraction by analyzing responses R c (l i ) = ln(m c(l i)) ln(max(m c(l k )), l k L c (1) R c (l i ) l i c 1 M c (l i ) 1 l i Wikipedia URL R c I R c Web Web [14] Web Web AND Web Web Web 3 Web AND N Web H H I N Web SoftBank 1 CaboCha [17] MeCab [18] v v c 1 c 2 l i l j p 1 p 2 F (p 1, p 2, v) F (p 1, p 2, v) = R c 1 (l i ) + R c2 (l j ) 2 F l i l j 1 c 1 c I(p 1, p 2, v) ( ) H(p1, p 2 ) I(p 1, p 2, v) = max, 1 N (2) F (p 1, p 2, v) (3) H(p 1, p 2) p 1 p 2 N Web (3) Web I 4. Web 4

5 4. 1 Wikipedia Web 5 Web (H) (A) (B) Web [14] 1 (CFR) 2 Mean Reciprocal Rank (MRR) 0 3 Discounted Cumulative Gain (DCG) Jarvelin [19] g(1) if i = 1 dcg(i) = dcg(i 1) + g(i) otherwise log c (i) h g(i) = a b if r(i) H if r(i) A if r(i) B r(i) i g(i) r(i) dcg(i) i [14], [20] (h, a, b) = (3, 2, 0) c = 2 0 CFR MRR DCG (4) (5) Table 2 Evaluation (all concept pairs) CFR CFR MRR MRR DCG DCG Table 3 Evaluation (concept pairs relation extracted between) CFR CFR MRR MRR DCG DCG Table 4 Analysis data in the experiment Web CFR MRR CFR 2 CFR MRR MRR 5

6 Table 5 5 An example of extracted relations Table 6 6 An example of extracted wrong relations A B Berryz A B NEWS A B ZARD A B A B A B A B A B A B A B A B A B A B A B SMAP A B DCG DCG DCG Web Web 50 Web Web A B MBS A B A B F A B A B A B A B A B A B A B SMAP NEWS MBS MBS MBS MBS MBS MBS MBS is-a a-part-of 6

7 is-a a-part-of 8! 5. Wikipedia Web Wikipedia Wikipedia [4], [5] Web 2 C( ) B( ) CORE [1] vol.20 no.6 pp Nov [2] vol.19 no.2 pp Mar [3] G.A. Miller, WordNet: A Lexical Database for English, Communications of the ACM (CACM), vol.38, no.11, pp.39 41, Nov [4] Wikipedia vol.47 no.10 pp Oct [5] K. Nakayama, T. Hara, and S. Nishio, Wikipedia Mining for An Association Web Thesaurus Construction, Proceedings of International Conference on Web Information Systems Engineering (WISE), pp , Dec [6] : Wikipedia vol.22 no.5 pp Sept [7] S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z.G. Ives, DBpedia: A Nucleus for a Web of Open Data, Proceedings of International Semantic Web Conference, Asian Semantic Web Conference (ISWC/ASWC), pp , Nov [8] F.M. Suchanek, G. Kasneci, and G. Weikum, YAGO: A Core of Semantic Knowledge, Proceedings of International Conference on World Wide Web (WWW), pp , May [9] F.M. Suchanek, G. Kasneci, and G. Weikum, YAGO: A Large Ontology from Wikipedia and WordNet, Journal of Web Semantics, vol.6, no.3, pp , Sept [10] Wikipedia 14 pp Mar [11] Wikipedia 18 Web SIG-SWO-A July [12] vol.12 no.2 pp Mar [13] D. Kawahara, and S. Kurohashi, Case Frame Compilation from the Web using High-Performance Computing, Proceedings of International Conference on Language Resources and Evaluation (LREC), May [14] D vol.j91-d no.3 pp Mar [15] Web : vol.47 no.sig19(tod32) pp Dec [16] K. Nakayama, T. Hara, and S. Nishio, A Thesaurus Construction Method from Large Scale Web Dictionaries, Proceedings of IEEE International Conference on Advanced Information Networking and Applications (AINA), pp , May [17] T. Kudo, and Y. Matsumoto, Fast Methods for Kernel- Based Text Analysis, Proceedings of Meeting on Association for Computational Linguistics (ACL), pp.24 31, July [18] T. Kudo, K. Yamamoto, and Y. Matsumoto, Applying Conditional Random Fields to Japanese Morphological Analysis, Proceedings of Conference on Empirical Methods in Natural Language Processing (EMNLP), pp , July [19] K. Järvelin, and J. Kekäläinen, IR Evaluation Methods for Retrieving Highly Relevant Documents, Proceedings of International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR), pp.41 48, July [20] K. Eguchi, Overview of the Topical Classification Task at NTCIR-4 WEB, Working Notes of the 4th NTCIR Meeting, Supplement, vol.1, no.48 55, June

THE SEMANTIC WEB AND IT`S APPLICATIONS

THE SEMANTIC WEB AND IT`S APPLICATIONS 15-16 September 2011, BULGARIA 1 Proceedings of the International Conference on Information Technologies (InfoTech-2011) 15-16 September 2011, Bulgaria THE SEMANTIC WEB AND IT`S APPLICATIONS Dimitar Vuldzhev

More information

INTERNATIONAL JOURNAL FOR TRENDS IN ENGINEERING & TECHNOLOGY VOLUME 3 ISSUE

INTERNATIONAL JOURNAL FOR TRENDS IN ENGINEERING & TECHNOLOGY VOLUME 3 ISSUE Enhancing Implicit Relations in Wikipedia Mining Using Object Relationship Technique G.Shanmugapriya 1 1 B.S Abdur Rahman University, Computer Science, sarushiya@gmail.com S.Raja shaik 2 2 B.S Abdur Rahman

More information

Discovering and Querying Hybrid Linked Data

Discovering and Querying Hybrid Linked Data Discovering and Querying Hybrid Linked Data Zareen Syed 1, Tim Finin 1, Muhammad Rahman 1, James Kukla 2, Jeehye Yun 2 1 University of Maryland Baltimore County 1000 Hilltop Circle, MD, USA 21250 zsyed@umbc.edu,

More information

Integrating FLOSS repositories on the Web

Integrating FLOSS repositories on the Web DERI DIGITAL ENTERPRISE RESEARCH INSTITUTE Integrating FLOSS repositories on the Web Aftab Iqbal Richard Cyganiak Michael Hausenblas DERI Technical Report 2012-12-10 December 2012 DERI Galway IDA Business

More information

ONTOLOGIES A short tutorial with references to YAGO Cosmina CROITORU

ONTOLOGIES A short tutorial with references to YAGO Cosmina CROITORU ONTOLOGIES p. 1/40 ONTOLOGIES A short tutorial with references to YAGO Cosmina CROITORU Unlocking the Secrets of the Past: Text Mining for Historical Documents Blockseminar, 21.2.-11.3.2011 ONTOLOGIES

More information

Subtask Mining from Search Query Logs for How-Knowledge Acceleration

Subtask Mining from Search Query Logs for How-Knowledge Acceleration Subtask Mining from Search Query Logs for How-Knowledge Acceleration Chung-Lun Kuo and Hsin-Hsi Chen Department of Computer Science and Information Engineering, National Taiwan University No. 1, Sec. 4,

More information

Cross-Lingual Concern Analysis from Multilingual Weblog Articles

Cross-Lingual Concern Analysis from Multilingual Weblog Articles Cross-Lingual Concern Analysis from Multilingual Weblog Articles Tomohiro Fukuhara RACE (Research into Artifacts), The University of Tokyo 5-1-5 Kashiwanoha, Kashiwa, Chiba JAPAN http://www.race.u-tokyo.ac.jp/~fukuhara/

More information

Towards the Integration of a Research Group Website into the Web of Data

Towards the Integration of a Research Group Website into the Web of Data Towards the Integration of a Research Group Website into the Web of Data Mikel Emaldi, David Buján, and Diego López-de-Ipiña Deusto Institute of Technology - DeustoTech, University of Deusto Avda. Universidades

More information

Monitoring Web Browsing Habits of User Using Web Log Analysis and Role-Based Web Accessing Control. Phudinan Singkhamfu, Parinya Suwanasrikham

Monitoring Web Browsing Habits of User Using Web Log Analysis and Role-Based Web Accessing Control. Phudinan Singkhamfu, Parinya Suwanasrikham Monitoring Web Browsing Habits of User Using Web Log Analysis and Role-Based Web Accessing Control Phudinan Singkhamfu, Parinya Suwanasrikham Chiang Mai University, Thailand 0659 The Asian Conference on

More information

Mining Signatures in Healthcare Data Based on Event Sequences and its Applications

Mining Signatures in Healthcare Data Based on Event Sequences and its Applications Mining Signatures in Healthcare Data Based on Event Sequences and its Applications Siddhanth Gokarapu 1, J. Laxmi Narayana 2 1 Student, Computer Science & Engineering-Department, JNTU Hyderabad India 1

More information

Intui2: A Prototype System for Question Answering over Linked Data

Intui2: A Prototype System for Question Answering over Linked Data Intui2: A Prototype System for Question Answering over Linked Data Corina Dima Seminar für Sprachwissenschaft, University of Tübingen, Wilhemstr. 19, 72074 Tübingen, Germany corina.dima@uni-tuebingen.de

More information

Probabilistic Semantic Similarity Measurements for Noisy Short Texts Using Wikipedia Entities

Probabilistic Semantic Similarity Measurements for Noisy Short Texts Using Wikipedia Entities Probabilistic Semantic Similarity Measurements for Noisy Short Texts Using Wikipedia Entities Masumi Shirakawa Kotaro Nakayama Takahiro Hara Shojiro Nishio Graduate School of Information Science and Technology,

More information

LiDDM: A Data Mining System for Linked Data

LiDDM: A Data Mining System for Linked Data LiDDM: A Data Mining System for Linked Data Venkata Narasimha Pavan Kappara Indian Institute of Information Technology Allahabad Allahabad, India kvnpavan@gmail.com Ryutaro Ichise National Institute of

More information

Automatic Mining of Internet Translation Reference Knowledge Based on Multiple Search Engines

Automatic Mining of Internet Translation Reference Knowledge Based on Multiple Search Engines , 22-24 October, 2014, San Francisco, USA Automatic Mining of Internet Translation Reference Knowledge Based on Multiple Search Engines Baosheng Yin, Wei Wang, Ruixue Lu, Yang Yang Abstract With the increasing

More information

Querying DBpedia Using HIVE-QL

Querying DBpedia Using HIVE-QL Querying DBpedia Using HIVE-QL AHMED SALAMA ISMAIL 1, HAYTHAM AL-FEEL 2, HODA M. O.MOKHTAR 3 Information Systems Department, Faculty of Computers and Information 1, 2 Fayoum University 3 Cairo University

More information

Semantic Content Management with Apache Stanbol

Semantic Content Management with Apache Stanbol Semantic Content Management with Apache Stanbol Ali Anil SINACI and Suat GONUL SRDC Software Research & Development and Consultancy Ltd., ODTU Teknokent Silikon Blok No:14, 06800 Ankara, Turkey {anil,suat}@srdc.com.tr

More information

Evaluation of Retrieval Systems

Evaluation of Retrieval Systems Performance Criteria Evaluation of Retrieval Systems 1 1. Expressiveness of query language Can query language capture information needs? 2. Quality of search results Relevance to users information needs

More information

Data Mining in Web Search Engine Optimization and User Assisted Rank Results

Data Mining in Web Search Engine Optimization and User Assisted Rank Results Data Mining in Web Search Engine Optimization and User Assisted Rank Results Minky Jindal Institute of Technology and Management Gurgaon 122017, Haryana, India Nisha kharb Institute of Technology and Management

More information

SEARCH ENGINE WITH PARALLEL PROCESSING AND INCREMENTAL K-MEANS FOR FAST SEARCH AND RETRIEVAL

SEARCH ENGINE WITH PARALLEL PROCESSING AND INCREMENTAL K-MEANS FOR FAST SEARCH AND RETRIEVAL SEARCH ENGINE WITH PARALLEL PROCESSING AND INCREMENTAL K-MEANS FOR FAST SEARCH AND RETRIEVAL Krishna Kiran Kattamuri 1 and Rupa Chiramdasu 2 Department of Computer Science Engineering, VVIT, Guntur, India

More information

Representing Specialized Events with FrameBase

Representing Specialized Events with FrameBase Representing Specialized Events with FrameBase Jacobo Rouces 1, Gerard de Melo 2, and Katja Hose 1 1 Aalborg University, Denmark jrg@es.aau.dk, khose@cs.aau.dk 2 Tsinghua University, China gdm@demelo.org

More information

Social Business Intelligence Text Search System

Social Business Intelligence Text Search System Social Business Intelligence Text Search System Sagar Ligade ME Computer Engineering. Pune Institute of Computer Technology Pune, India ABSTRACT Today the search engine plays the important role in the

More information

Semi-Supervised Learning for Blog Classification

Semi-Supervised Learning for Blog Classification Proceedings of the Twenty-Third AAAI Conference on Artificial Intelligence (2008) Semi-Supervised Learning for Blog Classification Daisuke Ikeda Department of Computational Intelligence and Systems Science,

More information

Cross-lingual Knowledge Linking Across Wiki Knowledge Bases

Cross-lingual Knowledge Linking Across Wiki Knowledge Bases Cross-lingual Knowledge Linking Across Wiki Knowledge Bases Zhichun Wang, Juanzi Li, Zhigang Wang, and Jie Tang Department of Computer Science and Technology Tsinghua University, Beijing, China {zcwang,

More information

Wikipedia and Web document based Query Translation and Expansion for Cross-language IR

Wikipedia and Web document based Query Translation and Expansion for Cross-language IR Wikipedia and Web document based Query Translation and Expansion for Cross-language IR Ling-Xiang Tang 1, Andrew Trotman 2, Shlomo Geva 1, Yue Xu 1 1Faculty of Science and Technology, Queensland University

More information

FUZZY CLUSTERING ANALYSIS OF DATA MINING: APPLICATION TO AN ACCIDENT MINING SYSTEM

FUZZY CLUSTERING ANALYSIS OF DATA MINING: APPLICATION TO AN ACCIDENT MINING SYSTEM International Journal of Innovative Computing, Information and Control ICIC International c 0 ISSN 34-48 Volume 8, Number 8, August 0 pp. 4 FUZZY CLUSTERING ANALYSIS OF DATA MINING: APPLICATION TO AN ACCIDENT

More information

Natural Language to Relational Query by Using Parsing Compiler

Natural Language to Relational Query by Using Parsing Compiler Available Online at www.ijcsmc.com International Journal of Computer Science and Mobile Computing A Monthly Journal of Computer Science and Information Technology IJCSMC, Vol. 4, Issue. 3, March 2015,

More information

Enhancing the relativity between Content, Title and Meta Tags Based on Term Frequency in Lexical and Semantic Aspects

Enhancing the relativity between Content, Title and Meta Tags Based on Term Frequency in Lexical and Semantic Aspects Enhancing the relativity between Content, Title and Meta Tags Based on Term Frequency in Lexical and Semantic Aspects Mohammad Farahmand, Abu Bakar MD Sultan, Masrah Azrifah Azmi Murad, Fatimah Sidi me@shahroozfarahmand.com

More information

SEMANTIC WEB BASED INFERENCE MODEL FOR LARGE SCALE ONTOLOGIES FROM BIG DATA

SEMANTIC WEB BASED INFERENCE MODEL FOR LARGE SCALE ONTOLOGIES FROM BIG DATA SEMANTIC WEB BASED INFERENCE MODEL FOR LARGE SCALE ONTOLOGIES FROM BIG DATA J.RAVI RAJESH PG Scholar Rajalakshmi engineering college Thandalam, Chennai. ravirajesh.j.2013.mecse@rajalakshmi.edu.in Mrs.

More information

Archivage du contenu éphémère du Web à l aide des flux Web *

Archivage du contenu éphémère du Web à l aide des flux Web * Archivage du contenu éphémère du Web à l aide des flux Web * Marilena Oita Pierre Senellart Résumé Cette proposition de démonstration concerne une application d archivage du contenu du Web à l aide des

More information

The Path is the Destination Enabling a New Search Paradigm with Linked Data

The Path is the Destination Enabling a New Search Paradigm with Linked Data The Path is the Destination Enabling a New Search Paradigm with Linked Data Jörg Waitelonis, Magnus Knuth, Lina Wolf, Johannes Hercher, and Harald Sack Hasso-Plattner-Institute Potsdam, Prof.-Dr.-Helmert-Str.

More information

Knowledge Continuous Integration Process (K-CIP)

Knowledge Continuous Integration Process (K-CIP) Knowledge Continuous Integration Process (K-CIP) Hala Skaf-Molli LINA Université de Nantes 2 rue de la Houssinière BP92208, F-44300 Nantes Cedex 3, France hala.skaf@univ-nantes.fr Gérôme Canals LORIA Université

More information

Developing Semantic Classifiers for Big Data

Developing Semantic Classifiers for Big Data Semantics for Big Data AAAI Technical Report FS-13-04 Developing Semantic Classifiers for Big Data Richard Scherl Department Computer Science & Stware Engineering Monmouth University West Long Branch,

More information

Exploiting Linked Open Data as Background Knowledge in Data Mining

Exploiting Linked Open Data as Background Knowledge in Data Mining Exploiting Linked Open Data as Background Knowledge in Data Mining Heiko Paulheim University of Mannheim, Germany Research Group Data and Web Science heiko@informatik.uni-mannheim.de Abstract. Many data

More information

Knowledge Continuous Integration Process (K-CIP)

Knowledge Continuous Integration Process (K-CIP) Knowledge Continuous Integration Process (K-CIP) Hala Skaf-Molli LINA Université de Nantes 2 rue de la Houssinière BP92208, F-44300 Nantes Cedex 3, France hala.skaf@univ-nantes.fr Gérôme Canals LORIA Université

More information

2-3 Automatic Construction Technology for Parallel Corpora

2-3 Automatic Construction Technology for Parallel Corpora 2-3 Automatic Construction Technology for Parallel Corpora We have aligned Japanese and English news articles and sentences, extracted from the Yomiuri and the Daily Yomiuri newspapers, to make a large

More information

Search Taxonomy. Web Search. Search Engine Optimization. Information Retrieval

Search Taxonomy. Web Search. Search Engine Optimization. Information Retrieval Information Retrieval INFO 4300 / CS 4300! Retrieval models Older models» Boolean retrieval» Vector Space model Probabilistic Models» BM25» Language models Web search» Learning to Rank Search Taxonomy!

More information

A Bi-Dimensional User Profile to Discover Unpopular Web Sources

A Bi-Dimensional User Profile to Discover Unpopular Web Sources A Bi-Dimensional User Profile to Discover Unpopular Web Sources Romain Noel Airbus Defense & Space Val-de-Reuil, France romain.noel@cassidian.com Laurent Vercouter St Etienne du Rouvray, France laurent.vercouter@insarouen.fr

More information

A Survey on Product Aspect Ranking Techniques

A Survey on Product Aspect Ranking Techniques A Survey on Product Aspect Ranking Techniques Ancy. J. S, Nisha. J.R P.G. Scholar, Dept. of C.S.E., Marian Engineering College, Kerala University, Trivandrum, India. Asst. Professor, Dept. of C.S.E., Marian

More information

Converging Web-Data and Database Data: Big - and Small Data via Linked Data

Converging Web-Data and Database Data: Big - and Small Data via Linked Data DBKDA/WEB Panel 2014, Chamonix, 24.04.2014 DBKDA/WEB Panel 2014, Chamonix, 24.04.2014 Reutlingen University Converging Web-Data and Database Data: Big - and Small Data via Linked Data Moderation: Fritz

More information

Survey Results: Requirements and Use Cases for Linguistic Linked Data

Survey Results: Requirements and Use Cases for Linguistic Linked Data Survey Results: Requirements and Use Cases for Linguistic Linked Data 1 Introduction This survey was conducted by the FP7 Project LIDER (http://www.lider-project.eu/) as input into the W3C Community Group

More information

POSBIOTM-NER: A Machine Learning Approach for. Bio-Named Entity Recognition

POSBIOTM-NER: A Machine Learning Approach for. Bio-Named Entity Recognition POSBIOTM-NER: A Machine Learning Approach for Bio-Named Entity Recognition Yu Song, Eunji Yi, Eunju Kim, Gary Geunbae Lee, Department of CSE, POSTECH, Pohang, Korea 790-784 Soo-Jun Park Bioinformatics

More information

Semantic annotation of requirements for automatic UML class diagram generation

Semantic annotation of requirements for automatic UML class diagram generation www.ijcsi.org 259 Semantic annotation of requirements for automatic UML class diagram generation Soumaya Amdouni 1, Wahiba Ben Abdessalem Karaa 2 and Sondes Bouabid 3 1 University of tunis High Institute

More information

On the Feasibility of Answer Suggestion for Advice-seeking Community Questions about Government Services

On the Feasibility of Answer Suggestion for Advice-seeking Community Questions about Government Services 21st International Congress on Modelling and Simulation, Gold Coast, Australia, 29 Nov to 4 Dec 2015 www.mssanz.org.au/modsim2015 On the Feasibility of Answer Suggestion for Advice-seeking Community Questions

More information

A Case Study of Question Answering in Automatic Tourism Service Packaging

A Case Study of Question Answering in Automatic Tourism Service Packaging BULGARIAN ACADEMY OF SCIENCES CYBERNETICS AND INFORMATION TECHNOLOGIES Volume 13, Special Issue Sofia 2013 Print ISSN: 1311-9702; Online ISSN: 1314-4081 DOI: 10.2478/cait-2013-0045 A Case Study of Question

More information

Automated FAQ Answering with Question-Specific Knowledge Representation for Web Self-Service

Automated FAQ Answering with Question-Specific Knowledge Representation for Web Self-Service HSI 2009 Catania, Italy, May 21-23, 2009 Automated FAQ Answering with Question-Specific Knowledge Representation for Web Self-Service Eriks Sneiders Dept. of Computer and Systems Sciences, Stockholm University

More information

Recent Topics of Research around the YAGO Knowledge Base

Recent Topics of Research around the YAGO Knowledge Base Recent Topics of Research around the YAGO Knowledge Base Antoine Amarilli 1, Luis Galárraga 1, Nicoleta Preda 2, and Fabian M. Suchanek 1 1 Télécom ParisTech, Paris, France 2 University of Versailles,

More information

Domain Classification of Technical Terms Using the Web

Domain Classification of Technical Terms Using the Web Systems and Computers in Japan, Vol. 38, No. 14, 2007 Translated from Denshi Joho Tsushin Gakkai Ronbunshi, Vol. J89-D, No. 11, November 2006, pp. 2470 2482 Domain Classification of Technical Terms Using

More information

Cross-Language Information Retrieval by Domain Restriction using Web Directory Structure

Cross-Language Information Retrieval by Domain Restriction using Web Directory Structure Cross-Language Information Retrieval by Domain Restriction using Web Directory Structure Fuminori Kimura Faculty of Culture and Information Science, Doshisha University 1 3 Miyakodani Tatara, Kyoutanabe-shi,

More information

LOCATION-AWARE MOBILE LEARNING OF SPATIAL ALGORITHMS

LOCATION-AWARE MOBILE LEARNING OF SPATIAL ALGORITHMS LOCATION-AWARE MOBILE LEARNING OF SPATIAL ALGORITHMS Ville Karavirta Department of Computer Science and Engineering, Aalto University PO. Box 15400, FI-00076 Aalto, FINLAND ABSTRACT Learning an algorithm

More information

Keywords: Information Retrieval, Vector Space Model, Database, Similarity Measure, Genetic Algorithm.

Keywords: Information Retrieval, Vector Space Model, Database, Similarity Measure, Genetic Algorithm. Volume 3, Issue 8, August 2013 ISSN: 2277 128X International Journal of Advanced Research in Computer Science and Software Engineering Research Paper Available online at: www.ijarcsse.com Effective Information

More information

ASKWIKI : SHALLOW SEMANTIC PROCESSING TO QUERY WIKIPEDIA. Felix Burkhardt and Jianshen Zhou. Deutsche Telekom Laboratories, Berlin, Germany

ASKWIKI : SHALLOW SEMANTIC PROCESSING TO QUERY WIKIPEDIA. Felix Burkhardt and Jianshen Zhou. Deutsche Telekom Laboratories, Berlin, Germany 20th European Signal Processing Conference (EUSIPCO 2012) Bucharest, Romania, August 27-31, 2012 ASKWIKI : SHALLOW SEMANTIC PROCESSING TO QUERY WIKIPEDIA Felix Burkhardt and Jianshen Zhou Deutsche Telekom

More information

Optimized Mobile Search Engine

Optimized Mobile Search Engine Optimized Mobile Search Engine E.Chaitanya 1, Dr.Sai Satyanarayana Reddy 2, O.Srinivasa Reddy 3 1 M.Tech, CSE, LBRCE, Mylavaram, 2 Professor, CSE, LBRCE, Mylavaram, India, 3 Asst.professor,CSE,LBRCE,Mylavaram,India.

More information

Wikidata. Semantic Web in Libraries December 2014. A Free Collaborative Knowledge Base. Markus Krötzsch TU Dresden

Wikidata. Semantic Web in Libraries December 2014. A Free Collaborative Knowledge Base. Markus Krötzsch TU Dresden Technische Universität Dresden Fakultät Informatik Wikidata A Free Collaborative Knowledge Base Markus Krötzsch TU Dresden Semantic Web in Libraries December 2014 Where is Wikipedia Going? Wikipedia in

More information

A MULTILINGUAL AND LOCATION EVALUATION OF SEARCH ENGINES FOR WEBSITES AND SEARCHED FOR KEYWORDS

A MULTILINGUAL AND LOCATION EVALUATION OF SEARCH ENGINES FOR WEBSITES AND SEARCHED FOR KEYWORDS A MULTILINGUAL AND LOCATION EVALUATION OF SEARCH ENGINES FOR WEBSITES AND SEARCHED FOR KEYWORDS Anas AlSobh Ahmed Al Oroud Mohammed N. Al-Kabi Izzat AlSmadi Yarmouk University Jordan ABSTRACT Search engines

More information

Sentiment analysis for news articles

Sentiment analysis for news articles Prashant Raina Sentiment analysis for news articles Wide range of applications in business and public policy Especially relevant given the popularity of online media Previous work Machine learning based

More information

Active Learning SVM for Blogs recommendation

Active Learning SVM for Blogs recommendation Active Learning SVM for Blogs recommendation Xin Guan Computer Science, George Mason University Ⅰ.Introduction In the DH Now website, they try to review a big amount of blogs and articles and find the

More information

An Approach to support Web Service Classification and Annotation

An Approach to support Web Service Classification and Annotation An Approach to support Web Service Classification and Annotation Marcello Bruno, Gerardo Canfora, Massimiliano Di Penta, and Rita Scognamiglio marcello.bruno@unisannio.it, canfora@unisannio.it, dipenta@unisannio.it,

More information

Using Data Mining for Mobile Communication Clustering and Characterization

Using Data Mining for Mobile Communication Clustering and Characterization Using Data Mining for Mobile Communication Clustering and Characterization A. Bascacov *, C. Cernazanu ** and M. Marcu ** * Lasting Software, Timisoara, Romania ** Politehnica University of Timisoara/Computer

More information

Automatic ontology-based User Profile Learning from heterogeneous Web Resources in a Big Data Context

Automatic ontology-based User Profile Learning from heterogeneous Web Resources in a Big Data Context Automatic ontology-based User Profile Learning from heterogeneous Web Resources in a Big Data Context Anett Hoppe Supervised by C. Nicolle and A. Roxin CheckSem Group, LE2I Université de Bourgogne Dijon,

More information

Curriculum Vitae Ruben Sipos

Curriculum Vitae Ruben Sipos Curriculum Vitae Ruben Sipos Mailing Address: 349 Gates Hall Cornell University Ithaca, NY 14853 USA Mobile Phone: +1 607-229-0872 Date of Birth: 8 October 1985 E-mail: rs@cs.cornell.edu Web: http://www.cs.cornell.edu/~rs/

More information

Linked Open Government Data Analytics

Linked Open Government Data Analytics Linked Open Government Data Analytics Evangelos Kalampokis 1,2, Efthimios Tambouris 1,2, Konstantinos Tarabanis 1,2 1 Information Technologies Institute, Centre for Research & Technology - Hellas, Greece

More information

Impact of Financial News Headline and Content to Market Sentiment

Impact of Financial News Headline and Content to Market Sentiment International Journal of Machine Learning and Computing, Vol. 4, No. 3, June 2014 Impact of Financial News Headline and Content to Market Sentiment Tan Li Im, Phang Wai San, Chin Kim On, Rayner Alfred,

More information

A Hybrid Approach for Multi-Faceted IR in Multimodal Domain

A Hybrid Approach for Multi-Faceted IR in Multimodal Domain A Hybrid Approach for Multi-Faceted IR in Multimodal Domain Serwah Sabetghadam, Ralf Bierig, Andreas Rauber Institute of Software Technology and Interactive Systems Vienna University of Technology Vienna,

More information

DBpedia: A Multilingual Cross-Domain Knowledge Base

DBpedia: A Multilingual Cross-Domain Knowledge Base DBpedia: A Multilingual Cross-Domain Knowledge Base Pablo N. Mendes 1, Max Jakob 2, Christian Bizer 1 1 Web Based Systems Group, Freie Universität Berlin, Germany 2 Neofonie GmbH, Berlin, Germany first.last@fu-berlin.de,

More information

Type Inference on Noisy RDF Data

Type Inference on Noisy RDF Data Type Inference on Noisy RDF Data Heiko Paulheim and Christian Bizer University of Mannheim, Germany Research Group Data and Web Science {heiko,chris}@informatik.uni-mannheim.de Abstract. Type information

More information

LDA Based Security in Personalized Web Search

LDA Based Security in Personalized Web Search LDA Based Security in Personalized Web Search R. Dhivya 1 / PG Scholar, B. Vinodhini 2 /Assistant Professor, S. Karthik 3 /Prof & Dean Department of Computer Science & Engineering SNS College of Technology

More information

Identifying Focus, Techniques and Domain of Scientific Papers

Identifying Focus, Techniques and Domain of Scientific Papers Identifying Focus, Techniques and Domain of Scientific Papers Sonal Gupta Department of Computer Science Stanford University Stanford, CA 94305 sonal@cs.stanford.edu Christopher D. Manning Department of

More information

SemWeB Semantic Web Browser Improving Browsing Experience with Semantic and Personalized Information and Hyperlinks

SemWeB Semantic Web Browser Improving Browsing Experience with Semantic and Personalized Information and Hyperlinks SemWeB Semantic Web Browser Improving Browsing Experience with Semantic and Personalized Information and Hyperlinks Melike Şah, Wendy Hall and David C De Roure Intelligence, Agents and Multimedia Group,

More information

Mining the Web of Linked Data with RapidMiner

Mining the Web of Linked Data with RapidMiner Mining the Web of Linked Data with RapidMiner Petar Ristoski, Christian Bizer, and Heiko Paulheim University of Mannheim, Germany Data and Web Science Group {petar.ristoski,heiko,chris}@informatik.uni-mannheim.de

More information

Sustaining Privacy Protection in Personalized Web Search with Temporal Behavior

Sustaining Privacy Protection in Personalized Web Search with Temporal Behavior Sustaining Privacy Protection in Personalized Web Search with Temporal Behavior N.Jagatheshwaran 1 R.Menaka 2 1 Final B.Tech (IT), jagatheshwaran.n@gmail.com, Velalar College of Engineering and Technology,

More information

External Semantic Annotation of Web-Databases

External Semantic Annotation of Web-Databases External Semantic Annotation of Web-Databases Benjamin Dönz, Dietmar Bruckner Institute of Computer Technology, Technical University of Vienna {doenz, bruckner}@ict.tuwien.ac.at Abstract-This paper presents

More information

LinksTo A Web2.0 System that Utilises Linked Data Principles to Link Related Resources Together

LinksTo A Web2.0 System that Utilises Linked Data Principles to Link Related Resources Together LinksTo A Web2.0 System that Utilises Linked Data Principles to Link Related Resources Together Owen Sacco 1 and Matthew Montebello 1, 1 University of Malta, Msida MSD 2080, Malta. {osac001, matthew.montebello}@um.edu.mt

More information

International Journal of Engineering Research-Online A Peer Reviewed International Journal Articles available online http://www.ijoer.

International Journal of Engineering Research-Online A Peer Reviewed International Journal Articles available online http://www.ijoer. REVIEW ARTICLE ISSN: 2321-7758 UPS EFFICIENT SEARCH ENGINE BASED ON WEB-SNIPPET HIERARCHICAL CLUSTERING MS.MANISHA DESHMUKH, PROF. UMESH KULKARNI Department of Computer Engineering, ARMIET, Department

More information

Word Taxonomy for On-line Visual Asset Management and Mining

Word Taxonomy for On-line Visual Asset Management and Mining Word Taxonomy for On-line Visual Asset Management and Mining Osmar R. Zaïane * Eli Hagen ** Jiawei Han ** * Department of Computing Science, University of Alberta, Canada, zaiane@cs.uaberta.ca ** School

More information

Sheeba J.I1, Vivekanandan K2

Sheeba J.I1, Vivekanandan K2 IMPROVED UNSUPERVISED FRAMEWORK FOR SOLVING SYNONYM, HOMONYM, HYPONYMY & POLYSEMY PROBLEMS FROM EXTRACTED KEYWORDS AND IDENTIFY TOPICS IN MEETING TRANSCRIPTS Sheeba J.I1, Vivekanandan K2 1 Assistant Professor,sheeba@pec.edu

More information

Profile Based Personalized Web Search and Download Blocker

Profile Based Personalized Web Search and Download Blocker Profile Based Personalized Web Search and Download Blocker 1 K.Sheeba, 2 G.Kalaiarasi Dhanalakshmi Srinivasan College of Engineering and Technology, Mamallapuram, Chennai, Tamil nadu, India Email: 1 sheebaoec@gmail.com,

More information

Evaluation of Bayesian Spam Filter and SVM Spam Filter

Evaluation of Bayesian Spam Filter and SVM Spam Filter Evaluation of Bayesian Spam Filter and SVM Spam Filter Ayahiko Niimi, Hirofumi Inomata, Masaki Miyamoto and Osamu Konishi School of Systems Information Science, Future University-Hakodate 116 2 Kamedanakano-cho,

More information

Linking Search Results, Bibliographical Ontologies and Linked Open Data Resources

Linking Search Results, Bibliographical Ontologies and Linked Open Data Resources Linking Search Results, Bibliographical Ontologies and Linked Open Data Resources Fabio Ricci, Javier Belmonte, Eliane Blumer, René Schneider Haute Ecole de Gestion de Genève, 7 route de Drize, CH-1227

More information

A Framework for Ontology-Based Knowledge Management System

A Framework for Ontology-Based Knowledge Management System A Framework for Ontology-Based Knowledge Management System Jiangning WU Institute of Systems Engineering, Dalian University of Technology, Dalian, 116024, China E-mail: jnwu@dlut.edu.cn Abstract Knowledge

More information

The Development of Multimedia-Multilingual Document Storage, Retrieval and Delivery System for E-Organization (STREDEO PROJECT)

The Development of Multimedia-Multilingual Document Storage, Retrieval and Delivery System for E-Organization (STREDEO PROJECT) The Development of Multimedia-Multilingual Storage, Retrieval and Delivery for E-Organization (STREDEO PROJECT) Asanee Kawtrakul, Kajornsak Julavittayanukool, Mukda Suktarachan, Patcharee Varasrai, Nathavit

More information

Using Semantic Data Mining for Classification Improvement and Knowledge Extraction

Using Semantic Data Mining for Classification Improvement and Knowledge Extraction Using Semantic Data Mining for Classification Improvement and Knowledge Extraction Fernando Benites and Elena Sapozhnikova University of Konstanz, 78464 Konstanz, Germany. Abstract. The objective of this

More information

Enriching the Crosslingual Link Structure of Wikipedia - A Classification-Based Approach -

Enriching the Crosslingual Link Structure of Wikipedia - A Classification-Based Approach - Enriching the Crosslingual Link Structure of Wikipedia - A Classification-Based Approach - Philipp Sorg and Philipp Cimiano Institute AIFB, University of Karlsruhe, D-76128 Karlsruhe, Germany {sorg,cimiano}@aifb.uni-karlsruhe.de

More information

Recovering Traceability Links between Requirements and Source Code using the Configuration Management Log *

Recovering Traceability Links between Requirements and Source Code using the Configuration Management Log * IEICE TRANS. FUNDAMENTALS/COMMUN./ELECTRON./INF. & SYST., VOL. E85-A/B/C/D, No. xx JANUARY 20xx 1 PAPER Recovering Traceability Links between Requirements and Source Code using the Configuration Management

More information

Query Recommendation employing Query Logs in Search Optimization

Query Recommendation employing Query Logs in Search Optimization 1917 Query Recommendation employing Query Logs in Search Optimization Neha Singh Department of Computer Science, Shri Siddhi Vinayak Group of Institutions, Bareilly Email: singh26.neha@gmail.com Dr Manish

More information

Creating an RDF Graph from a Relational Database Using SPARQL

Creating an RDF Graph from a Relational Database Using SPARQL Creating an RDF Graph from a Relational Database Using SPARQL Ayoub Oudani, Mohamed Bahaj*, Ilias Cherti Department of Mathematics and Informatics, University Hassan I, FSTS, Settat, Morocco. * Corresponding

More information

Analysis One Code Desc. Transaction Amount. Fiscal Period

Analysis One Code Desc. Transaction Amount. Fiscal Period Analysis One Code Desc Transaction Amount Fiscal Period 57.63 Oct-12 12.13 Oct-12-38.90 Oct-12-773.00 Oct-12-800.00 Oct-12-187.00 Oct-12-82.00 Oct-12-82.00 Oct-12-110.00 Oct-12-1115.25 Oct-12-71.00 Oct-12-41.00

More information

ONLINE RESUME PARSING SYSTEM USING TEXT ANALYTICS

ONLINE RESUME PARSING SYSTEM USING TEXT ANALYTICS ONLINE RESUME PARSING SYSTEM USING TEXT ANALYTICS Divyanshu Chandola 1, Aditya Garg 2, Ankit Maurya 3, Amit Kushwaha 4 1 Student, Department of Information Technology, ABES Engineering College, Uttar Pradesh,

More information

Harvesting and Structuring Social Data in Music Information Retrieval

Harvesting and Structuring Social Data in Music Information Retrieval Harvesting and Structuring Social Data in Music Information Retrieval Sergio Oramas Music Technology Group Universitat Pompeu Fabra, Barcelona, Spain sergio.oramas@upf.edu Abstract. An exponentially growing

More information

SINAI at WEPS-3: Online Reputation Management

SINAI at WEPS-3: Online Reputation Management SINAI at WEPS-3: Online Reputation Management M.A. García-Cumbreras, M. García-Vega F. Martínez-Santiago and J.M. Peréa-Ortega University of Jaén. Departamento de Informática Grupo Sistemas Inteligentes

More information

Automatic Annotation Wrapper Generation and Mining Web Database Search Result

Automatic Annotation Wrapper Generation and Mining Web Database Search Result Automatic Annotation Wrapper Generation and Mining Web Database Search Result V.Yogam 1, K.Umamaheswari 2 1 PG student, ME Software Engineering, Anna University (BIT campus), Trichy, Tamil nadu, India

More information

Big Data in The Web. Agenda. Big Data Asking the Right Questions Wisdom of Crowds in the Web The Long Tail Issues and Examples Concluding Remarks

Big Data in The Web. Agenda. Big Data Asking the Right Questions Wisdom of Crowds in the Web The Long Tail Issues and Examples Concluding Remarks Big Data in The Web Ricardo Baeza-Yates Yahoo! Labs Barcelona & Santiago de Chile Agenda Big Data Asking the Right Questions Wisdom of Crowds in the Web The Long Tail Issues and Examples Concluding Remarks

More information

Site-Specific versus General Purpose Web Search Engines: A Comparative Evaluation

Site-Specific versus General Purpose Web Search Engines: A Comparative Evaluation Panhellenic Conference on Informatics Site-Specific versus General Purpose Web Search Engines: A Comparative Evaluation G. Atsaros, D. Spinellis, P. Louridas Department of Management Science and Technology

More information

Building a Spanish MMTx by using Automatic Translation and Biomedical Ontologies

Building a Spanish MMTx by using Automatic Translation and Biomedical Ontologies Building a Spanish MMTx by using Automatic Translation and Biomedical Ontologies Francisco Carrero 1, José Carlos Cortizo 1,2, José María Gómez 3 1 Universidad Europea de Madrid, C/Tajo s/n, Villaviciosa

More information

Statistical Analyses of Named Entity Disambiguation Benchmarks

Statistical Analyses of Named Entity Disambiguation Benchmarks Statistical Analyses of Named Entity Disambiguation Benchmarks Nadine Steinmetz, Magnus Knuth, and Harald Sack Hasso Plattner Institute for Software Systems Engineering, Potsdam, Germany, firstname.lastname@hpi.uni-potsdam.de

More information

An Information Retrieval System for Expert and Consumer Users

An Information Retrieval System for Expert and Consumer Users An Information Retrieval System for Expert and Consumer Users Rena Peraki, Euripides G.M. Petrakis, Angelos Hliaoutakis Department of Electronic and Computer Engineering Technical University of Crete (TUC)

More information

Kybots, knowledge yielding robots German Rigau IXA group, UPV/EHU http://ixa.si.ehu.es

Kybots, knowledge yielding robots German Rigau IXA group, UPV/EHU http://ixa.si.ehu.es KYOTO () Intelligent Content and Semantics Knowledge Yielding Ontologies for Transition-Based Organization http://www.kyoto-project.eu/ Kybots, knowledge yielding robots German Rigau IXA group, UPV/EHU

More information

How To Use Data Mining For Knowledge Management In Technology Enhanced Learning

How To Use Data Mining For Knowledge Management In Technology Enhanced Learning Proceedings of the 6th WSEAS International Conference on Applications of Electrical Engineering, Istanbul, Turkey, May 27-29, 2007 115 Data Mining for Knowledge Management in Technology Enhanced Learning

More information

Enhancing Requirement Traceability Link Using User's Updating Activity

Enhancing Requirement Traceability Link Using User's Updating Activity ISSN (Online) : 2319-8753 ISSN (Print) : 2347-6710 International Journal of Innovative Research in Science, Engineering and Technology Volume 3, Special Issue 3, March 2014 2014 International Conference

More information

Performance Evaluation Techniques for an Automatic Question Answering System

Performance Evaluation Techniques for an Automatic Question Answering System Performance Evaluation Techniques for an Automatic Question Answering System Tilani Gunawardena, Nishara Pathirana, Medhavi Lokuhetti, Roshan Ragel, and Sampath Deegalla Abstract Automatic question answering

More information