skip to main content
article

Detecting context-differentiating terms using competitive learning

Published:01 September 2003Publication History
Skip Abstract Section

Abstract

Personal information agents monitor ongoing user information accesses in order to provide users with context-relevant information. Providing the needed information requires effective methods for identifying the user's task context, based on available information. For user browsing tasks, one approach to context identification is to extract context-determining terms from the documents that the user consults. The thesis of this article is (1) that term extraction for personal information agents can be done by learning terms whose occurrence frequencies have a large variance over time, (2) that indexing and retrieval based on these terms can be at least as effective as standard information retrieval techniques, and (3) that this information can be learned without comprehensive corpus analysis, making it suitable for use in personal information retrieval.We have developed an unsupervised term extraction algorithm, WordSieve, that learns individualized context-differentiating terms for document indexing and retrieval. This article presents a new version of WordSieve, compares its design and performance to our initial approach, and assesses its effectiveness for a controlled personal information retrieval task, compared to three common indexing techniques requiring statistics about the global corpus. In the experiments, the new version of WordSieve generates task-relevant indices of comparable or better quality to common indexing techniques, using only local information.

References

  1. Gediminas Adomavicius and Alexander Tuzhilin. Using data mining methods to build customer profiles. Computer, 34(2):74--82, February 2001.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  2. Ricardo Baeza-Yates and Berthier Ribeiro-Neto. Modern Information Retrieval. ACM Press, 1999.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  3. Marko Balabanović and Yoav Shoham. Learning information retrieval agents: Experiments with automated web browsing. In Proceedings of the AAAI Spring Symposium on Information Gathering from Heterogeneous, Distributed-Resources, March 1995.]]Google ScholarGoogle Scholar
  4. Travis Bauer and David Leake. Real time user context modeling for information retrieval agents. In Tenth International Conference on Information and Knowledge Management, pages 568--570. ACM Press, 2001.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  5. Travis Bauer and David Leake. Wordsieve: A method for real-time context extraction. In Modeling and Using Context: Proceedings of the Third International and Interdisciplinary Conference, Context 2001, pages 30--44.Springer-Verlag, 2001.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  6. Travis Bauer and David Leake. Calvin: A multi-agent persoanl information retrieval system. In Agent Oriented Information Systems 2002: Proceedings of the Fourth International Bi-Conference Workshop, 2002.]]Google ScholarGoogle Scholar
  7. Travis Bauer and David Leake. Using document access sequences to recommend customized information. IEEE Intelligent Systems, 17(6):27--32, Nov/Dec 2002.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  8. J. Budzik, K. Hammond, and L. Birnbaum. Information access in context. In Knowledge based systems, 2001.]]Google ScholarGoogle Scholar
  9. T. Joachims, D. Freitag, and T. Mitchell. Webwatcher: A tour guide for the world wide web. In Proceedings of IJCA197, August 1997.]]Google ScholarGoogle Scholar
  10. David B. Leake and Ryan Scherle. Towards context-based search engine selection. In Proceedings on the International Conference on Intelligent User Interfaces, pages 109--112, Santa Fe, NM, Jan 2001.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  11. Seng Wai Loke, Andrew Davison, and Leon Sterling. CIFI: An intelligent agent for citation finding on the world-wide web. In Pacific Rim International Conference on Artificial Intelligence, pages 580--591, 1996.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  12. U. Manber, M. Smith, and B. Gopal. Webglimpse: Combining browsing and searching. In Proceedings of 1997 Usenix Technical Conference, 1997.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  13. M. Pazzani, J. Muramatsu, and D. Billsus. Syskill & webert: Identifying interesting web sites. In Proceedings of the National Conference on Artificial Intelligence, Portland, OR, 1996.]]Google ScholarGoogle ScholarDigital LibraryDigital Library
  14. M. Pazzani, L. Nguyen, and S. Mantik. Learning from hotlists and coldlists: towards a www information filtering and seeking agent. In Proceedings of AI Tools Conference, Washington, DC, 1995.]] Google ScholarGoogle ScholarDigital LibraryDigital Library
  15. Murray R. Spiegel. Mathematical Handbook of Formulas and Tables. Shaum's Outline Series in Mathematics. McGraw-Hill Book Company, 1968.]]Google ScholarGoogle Scholar
  16. Ahmad M. Ahmad Wasfi. Collecting user access patterns for building user profiles and collaborative filtering. In Proceedings of the 4th international conference on Intelligent user interfaces, pages 57--64. ACM Press, 1999.]] Google ScholarGoogle ScholarDigital LibraryDigital Library

Index Terms

  1. Detecting context-differentiating terms using competitive learning
    Index terms have been assigned to the content through auto-classification.

    Recommendations

    Comments

    Login options

    Check if you have access through your login credentials or your institution to get full access on this article.

    Sign in

    Full Access

    • Published in

      cover image ACM SIGIR Forum
      ACM SIGIR Forum  Volume 37, Issue 2
      Fall 2003
      76 pages
      ISSN:0163-5840
      DOI:10.1145/959258
      Issue’s Table of Contents

      Copyright © 2003 Authors

      Publisher

      Association for Computing Machinery

      New York, NY, United States

      Publication History

      • Published: 1 September 2003

      Check for updates

      Qualifiers

      • article

    PDF Format

    View or Download as a PDF file.

    PDF

    eReader

    View online with eReader.

    eReader