<?xml version="1.0" encoding="UTF-8"?>
<?xml-stylesheet type="text/xsl" href="/oai-pmh.xsl"?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-09-18T19:53:02Z</responseDate>
  <request identifier="oai:www.ideals.illinois.edu:2142/16897" metadataPrefix="etdms" verb="GetRecord">https://www.ideals.illinois.edu/oai-pmh</request>
  <GetRecord>
    <record>
      <header>
        <identifier>oai:www.ideals.illinois.edu:2142/16897</identifier>
        <datestamp>2023-07-10</datestamp>
        <setSpec>col_2142_5131</setSpec>
        <setSpec>col_2142_10761</setSpec>
        <setSpec>com_2142_5130</setSpec>
        <setSpec>com_2142_10755</setSpec>
        <setSpec>com_2142_234</setSpec>
      </header>
      <metadata>
        <thesis xmlns="http://www.ndltd.org/standards/metadata/etdms/1.1/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.ndltd.org/standards/metadata/etdms/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdms11.xsd http://purl.org/dc/elements/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdmsdc.xsd">
          <dc:contributor>Zhai, ChengXiang</dc:contributor>
          <dc:creator>Hwang, Hsiang-Yeh</dc:creator>
          <dc:date>2010-08-20T18:01:10Z</dc:date>
          <dc:date>2010-08-20T18:01:10Z</dc:date>
          <dc:date>2010-08-20T18:01:10Z</dc:date>
          <dc:date>2010-08</dc:date>
          <dc:description>In this thesis, we study the problem of fetching informative sentences from multi-product documents. A multi-product document is defined as a document that mentions about multiple products, which is often for comparison purpose. An informative sentence is defined as a sentence that provides the characteristics of the product. we propose to use Probabilistic Latent Semantic Analysis (PLSA) to mine informative sentences given a single multi-product document. By applying PLSA in a multi-product document, it simultaneously solves three problems regarding to multi-product document: 1. Separate the document based on the products. 2. Fetch the informative key words for each product. 3. Fetch the informative sentences. The proposed method can mine the product information from a single multi-product document, which is quite different from previous works in opinion mining.
Experiment results show that the high probability words in the word distribution of each product do discover the characteristics of the product. The results also reveal that the sentences which contain more these keywords are more informative. Practical applications of this method are: 1. Product comparison. 2. Facilitate user in reading multi-product document, which is to apply data mining in human computer interaction (HCI).</dc:description>
          <dc:description>Item withdrawn by Mark Zulauf (zulauf@illinois.edu) on 2010-07-20T13:33:41Z
Item was in collections:
University of Illinois Theses &amp; Dissertations (ID: 1)
No. of bitstreams: 1
Hsiang-Yeh_Hwang.pdf: 2735829 bytes, checksum: 6596d844cc5ec2fe1384bcfd7965db80 (MD5)</dc:description>
          <dc:description>Made available in DSpace on 2010-08-20T18:01:10Z (GMT). No. of bitstreams: 3
Hsiang-Yeh_Hwang.pdf: 2735829 bytes, checksum: 6596d844cc5ec2fe1384bcfd7965db80 (MD5)
Hwang_Hsiang-Yeh.pdf: 2735824 bytes, checksum: 1b3935938aaf6a38430880d6832ef66d (MD5)
license.txt: 4061 bytes, checksum: a793f64dbfe95c45162b22c5887d7e25 (MD5)</dc:description>
          <dc:identifier>http://hdl.handle.net/2142/16897</dc:identifier>
          <dc:language>en</dc:language>
          <dc:rights>Copyright 2010 Hsiang-Yeh Hwang</dc:rights>
          <dc:subject>multi-product document</dc:subject>
          <dc:subject>single-product document</dc:subject>
          <dc:subject>informative sentences</dc:subject>
          <dc:subject>informative words (ie. keywords)</dc:subject>
          <dc:subject>product name entity</dc:subject>
          <dc:title>Mining informative sentences in multi-product documents with PLSA</dc:title>
          <degree>
            <department>Computer Science</department>
            <departmentCode>1434</departmentCode>
            <discipline>Computer Science</discipline>
            <disciplineCode>0112</disciplineCode>
            <grantor>University of Illinois at Urbana-Champaign</grantor>
            <level>Thesis</level>
            <name>M.S.</name>
            <program>MS:Computer Science -UIUC</program>
            <programCode>10KS0112MS</programCode>
          </degree>
        </thesis>
      </metadata>
    </record>
  </GetRecord>
</OAI-PMH>
