Nothing Special   »   [go: up one dir, main page]

skip to main content
article

Detecting context-differentiating terms using competitive learning

Published: 01 September 2003 Publication History

Abstract

Personal information agents monitor ongoing user information accesses in order to provide users with context-relevant information. Providing the needed information requires effective methods for identifying the user's task context, based on available information. For user browsing tasks, one approach to context identification is to extract context-determining terms from the documents that the user consults. The thesis of this article is (1) that term extraction for personal information agents can be done by learning terms whose occurrence frequencies have a large variance over time, (2) that indexing and retrieval based on these terms can be at least as effective as standard information retrieval techniques, and (3) that this information can be learned without comprehensive corpus analysis, making it suitable for use in personal information retrieval.We have developed an unsupervised term extraction algorithm, WordSieve, that learns individualized context-differentiating terms for document indexing and retrieval. This article presents a new version of WordSieve, compares its design and performance to our initial approach, and assesses its effectiveness for a controlled personal information retrieval task, compared to three common indexing techniques requiring statistics about the global corpus. In the experiments, the new version of WordSieve generates task-relevant indices of comparable or better quality to common indexing techniques, using only local information.

References

[1]
Gediminas Adomavicius and Alexander Tuzhilin. Using data mining methods to build customer profiles. Computer, 34(2):74--82, February 2001.]]
[2]
Ricardo Baeza-Yates and Berthier Ribeiro-Neto. Modern Information Retrieval. ACM Press, 1999.]]
[3]
Marko Balabanović and Yoav Shoham. Learning information retrieval agents: Experiments with automated web browsing. In Proceedings of the AAAI Spring Symposium on Information Gathering from Heterogeneous, Distributed-Resources, March 1995.]]
[4]
Travis Bauer and David Leake. Real time user context modeling for information retrieval agents. In Tenth International Conference on Information and Knowledge Management, pages 568--570. ACM Press, 2001.]]
[5]
Travis Bauer and David Leake. Wordsieve: A method for real-time context extraction. In Modeling and Using Context: Proceedings of the Third International and Interdisciplinary Conference, Context 2001, pages 30--44.Springer-Verlag, 2001.]]
[6]
Travis Bauer and David Leake. Calvin: A multi-agent persoanl information retrieval system. In Agent Oriented Information Systems 2002: Proceedings of the Fourth International Bi-Conference Workshop, 2002.]]
[7]
Travis Bauer and David Leake. Using document access sequences to recommend customized information. IEEE Intelligent Systems, 17(6):27--32, Nov/Dec 2002.]]
[8]
J. Budzik, K. Hammond, and L. Birnbaum. Information access in context. In Knowledge based systems, 2001.]]
[9]
T. Joachims, D. Freitag, and T. Mitchell. Webwatcher: A tour guide for the world wide web. In Proceedings of IJCA197, August 1997.]]
[10]
David B. Leake and Ryan Scherle. Towards context-based search engine selection. In Proceedings on the International Conference on Intelligent User Interfaces, pages 109--112, Santa Fe, NM, Jan 2001.]]
[11]
Seng Wai Loke, Andrew Davison, and Leon Sterling. CIFI: An intelligent agent for citation finding on the world-wide web. In Pacific Rim International Conference on Artificial Intelligence, pages 580--591, 1996.]]
[12]
U. Manber, M. Smith, and B. Gopal. Webglimpse: Combining browsing and searching. In Proceedings of 1997 Usenix Technical Conference, 1997.]]
[13]
M. Pazzani, J. Muramatsu, and D. Billsus. Syskill & webert: Identifying interesting web sites. In Proceedings of the National Conference on Artificial Intelligence, Portland, OR, 1996.]]
[14]
M. Pazzani, L. Nguyen, and S. Mantik. Learning from hotlists and coldlists: towards a www information filtering and seeking agent. In Proceedings of AI Tools Conference, Washington, DC, 1995.]]
[15]
Murray R. Spiegel. Mathematical Handbook of Formulas and Tables. Shaum's Outline Series in Mathematics. McGraw-Hill Book Company, 1968.]]
[16]
Ahmad M. Ahmad Wasfi. Collecting user access patterns for building user profiles and collaborative filtering. In Proceedings of the 4th international conference on Intelligent user interfaces, pages 57--64. ACM Press, 1999.]]

Cited By

View all
  • (2009)Describing and predicting information‐seeking behavior on the WebJournal of the American Society for Information Science and Technology10.1002/asi.2103560:4(679-693)Online publication date: 2-Feb-2009

Index Terms

  1. Detecting context-differentiating terms using competitive learning

    Recommendations

    Comments

    Please enable JavaScript to view thecomments powered by Disqus.

    Information & Contributors

    Information

    Published In

    cover image ACM SIGIR Forum
    ACM SIGIR Forum  Volume 37, Issue 2
    Fall 2003
    76 pages
    ISSN:0163-5840
    DOI:10.1145/959258
    Issue’s Table of Contents

    Publisher

    Association for Computing Machinery

    New York, NY, United States

    Publication History

    Published: 01 September 2003
    Published in SIGIR Volume 37, Issue 2

    Check for updates

    Qualifiers

    • Article

    Contributors

    Other Metrics

    Bibliometrics & Citations

    Bibliometrics

    Article Metrics

    • Downloads (Last 12 months)0
    • Downloads (Last 6 weeks)0
    Reflects downloads up to 18 Nov 2024

    Other Metrics

    Citations

    Cited By

    View all
    • (2009)Describing and predicting information‐seeking behavior on the WebJournal of the American Society for Information Science and Technology10.1002/asi.2103560:4(679-693)Online publication date: 2-Feb-2009

    View Options

    Login options

    View options

    PDF

    View or Download as a PDF file.

    PDF

    eReader

    View online with eReader.

    eReader

    Media

    Figures

    Other

    Tables

    Share

    Share

    Share this Publication link

    Share on social media