Web bilgi kaynaklarının konu-merkezli sorgulanması
2001
0 views
0 downloads
Advisor: Doç. Dr. Özgür Ulusoy
Abstract (EN)
ABSTRACT TOPIC-CENTRIC QUERYING OF WEB RESOURCES Ismail Sengör Altmgövde M.S. in Computer Engineering Supervisor: Assoc. Prof. Dr. Özgür Ulusoy September, 2001 As the world wide web (WWW) has evolved to be almost the largest source of information that is known by human beings, locating relevant information on the web in a reasonably short time has become a major struggle. High qual ity indices and (sometimes specialized) search engines that employ information retrieval techniques are widely used for keyword based searches, and a number of web query languages have also been developed, mostly for research purposes. However, most of the keyword-based approaches are vulnerable to the noise on the web, leading to unqualified results with lots of irrelevant documents; whereas the web-query languages lack the speed or generality to be used in practical cases. In this thesis, we make use of metadata (along with some XML-based stan dards) to characterize the web resource domains, and to provide sophisticated querying features with high-quality results and a reasonably fast response time. We propose a "web information space" metadata model for web information re sources, and a query language SQL-TC (Topic-Centric SQL) to query the model. The web information space model is composed of web-based information resources (XML or HTML documents on the web), expert advice repositories (domain ex pert specified metadata for information resources), and personalized information about users (user profiles and preferences, as XML documents). Expert advice is specified using topics and relationships among topics (called metalinks) in a particular domain of interest, along the lines of the recently proposed topic maps. Experts also attach importance values to topics and metalinks that they spec ify, and link them to actual information resources on the web whenever possible, creating a semantic index over the resources. User profiles keep track of user knowledge and navigation history in terms of these topics and their (visited) sources, whereas user preferences declare users' attitudes and confidence for the mIV choices of particular experts. The query language SQL-TC makes use of the metadata information provided in expert advice repositories and embedded in information resources, and employs user preferences to further refine the query output. Query output objects/tuples are ranked with respect to the (expert-judged and user-preference- revised) im portance values of requested topics/metalinks, and the query output is limited by either top ra-ranked objects/tuples, or objects/tuples with importance values above a given threshold, or both. Therefore, the query output of SQL-TC is expected to produce highly relevant and semantically related responses to user queries within short amounts of time. Keywords: metadata, XML, Topic Maps, web data modeling, web querying, se mantic indexing, user profile.
Author
Dr. İsmail Şengör Altıngövde
How to Cite
İsmail Şengör Altıngövde (Master Thesis). Web bilgi kaynaklarının konu-merkezli sorgulanması, 2001, Bilkent University.
License
Tüm Hakları Saklıdır
This work is shared under the specified license terms.
More theses from Bilkent University
- Geç Antik Çağ'da Aşağı Tuna: Histria örneği(2023)
- Petrol fiyatları ve getiri eğrisi(2024)
- Sözle yönlendirme üzerine makaleler(2014)
- İletişim ağları ve sağlık uygulamaları için çok kollu haydut algoritmaları(2022)
- Türk Anayasa Mahkemesinin içtihatları ışığında karşılaştırmalı anayasal mutluluk(2023)
- Doğrusal karbon zincirlerinin yoğunluk fonksiyoneli teorisi ile incelenmesi(2023)
