Finding Web Document Associations Using Frequent Pairs of Adjacent Words

Yong-Jin Tee, Jason; Soon, Lay-Ki; Ranaivo-Malançon, Bali

doi:10.1007/978-3-642-32826-8_39

Jason Yong-Jin Tee³,
Lay-Ki Soon³ &
Bali Ranaivo-Malançon³

Part of the book series: Communications in Computer and Information Science ((CCIS,volume 295))

Included in the following conference series:

Knowledge Technology Week

1011 Accesses

Abstract

This paper presents an approach to find associations between Web documents using collocated word pairs. Given two Web documents which are connected via a hyperlink, we attempt to find the contextual association of these two Web pages by using collocations of word pairs from a statistical point of view. Our preliminary experimental results show that our approach is able to extract fairly coherent word pairs to derive associations between hyperlinked Web documents.

This is a preview of subscription content, log in via an institution to check access.

Access this chapter

Log in via an institution

Subscribe and save

Springer+ Basic

$34.99 /Month

Get 10 units per month
Download Article/Chapter or eBook
1 Unit = 1 Article or 1 Chapter
Cancel anytime

Buy Now

Chapter: USD 29.95; Price excludes VAT (USA)

eBook: USD 39.99; Price excludes VAT (USA)

Softcover Book: USD 54.99; Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Purchases are for personal use only

Institutional subscriptions

Preview

Unable to display preview. Download preview PDF.

Correlating Words - Approaches and Applications

Large Scale Semantic Relation Discovery: Toward Establishing the Missing Link Between Wikipedia and Semantic Network

Semantic association computation: a comprehensive survey

Article 20 November 2019

References

Ahonen-Myka, H.: Discovery of Frequent Word Sequences in Text. In: Hand, D.J., Adams, N.M., Bolton, R.J. (eds.) Pattern Detection and Discovery. LNCS (LNAI), vol. 2447, pp. 180–328. Springer, Heidelberg (2002)
Chapter Google Scholar
Qureshi, M., Younus, A., Rojas, F.: Analyzing Web Crawler as Feed Forward Engine for Efficient Solution to Search Problem in the Minimum Amount of Time through a Distributed Framework. In: Proceedings of 1st International Conference on Information Science and Applications, Seoul, Korea (2010)
Google Scholar

Download references

Author information

Authors and Affiliations

Faculty of Information Technology, Multimedia University, Cyberjaya, Selangor, Malaysia
Jason Yong-Jin Tee, Lay-Ki Soon & Bali Ranaivo-Malançon

Authors

Jason Yong-Jin Tee
View author publications
You can also search for this author in PubMed Google Scholar
Lay-Ki Soon
View author publications
You can also search for this author in PubMed Google Scholar
Bali Ranaivo-Malançon
View author publications
You can also search for this author in PubMed Google Scholar

Editor information

Editors and Affiliations

MIMOS Berhad, Knowledge Technology, Technology Park Malaysia, 57000, Kuala Lumpur, Malaysia
Dickson Lukose
Department of Systems and Network,College of Information Technology, Jalan IKRAM-UNITEN, Universiti Tenaga Nasional, 43000, Kajang, Malaysia
Abdul Rahim Ahmad & Azizah Suliman &

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Yong-Jin Tee, J., Soon, LK., Ranaivo-Malançon, B. (2012). Finding Web Document Associations Using Frequent Pairs of Adjacent Words. In: Lukose, D., Ahmad, A.R., Suliman, A. (eds) Knowledge Technology. KTW 2011. Communications in Computer and Information Science, vol 295. Springer, Berlin, Heidelberg. https://doi.org/10.1007/978-3-642-32826-8_39

Download citation

DOI: https://doi.org/10.1007/978-3-642-32826-8_39
Publisher Name: Springer, Berlin, Heidelberg
Print ISBN: 978-3-642-32825-1
Online ISBN: 978-3-642-32826-8
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics