Skip to main navigation Skip to search Skip to main content

Structure- and Content-Based Retrieval for XML Documents

    Research output: Contribution to conferenceChapterpeer-review

    Abstract

    The XML was proposed as a standard markup language to make Web documents in 1996 (Extensible Markup Language, 2000). It has as good an expressive power as SGML and is easy to use like HTML. Recently, it has been common for users to acquire through the Web a variety of multimedia documents written by XML. Meanwhile, because the number of XML documents is dramatically increasing, it is difficult to reach a specific XML document required by users. Moreover, an XML document not only has a logical and hierarchical structure in common, but also contains its multimedia data, such as image and video. Thus, it is necessary to retrieve XML documents based on both document structure and image content. For supporting the structure-based retrieval, it is necessary to design four efficient index structures, that is, keyword, structure, element, and attribute index, by indexing XML documents using a basic element unit. For supporting the content-based retrieval, it is necessary to design a highdimensional index structure so as to store and retrieve both color and shape feature vectors efficiently.

    Original languageEnglish
    Title of host publicationEncyclopedia of Information Science and Technology
    Subtitle of host publicationVolume I
    PublisherIGI Global
    Pages2662-2664
    Number of pages3
    Volume1
    ISBN (Electronic)9781591407942
    ISBN (Print)9781591405535
    DOIs
    StatePublished - 2005.01.1

    Fingerprint

    Dive into the research topics of 'Structure- and Content-Based Retrieval for XML Documents'. Together they form a unique fingerprint.

    Cite this