<?xml version="1.0" encoding="UTF-8"?>
<?xml-stylesheet type="text/xsl" href="/oai-pmh.xsl"?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-09-19T08:51:40Z</responseDate>
  <request identifier="oai:www.ideals.illinois.edu:2142/26317" metadataPrefix="etdms" verb="GetRecord">https://www.ideals.illinois.edu/oai-pmh</request>
  <GetRecord>
    <record>
      <header>
        <identifier>oai:www.ideals.illinois.edu:2142/26317</identifier>
        <datestamp>2023-07-10</datestamp>
        <setSpec>col_2142_5131</setSpec>
        <setSpec>col_2142_8888</setSpec>
        <setSpec>com_2142_5130</setSpec>
        <setSpec>com_2142_8887</setSpec>
        <setSpec>com_2142_234</setSpec>
      </header>
      <metadata>
        <thesis xmlns="http://www.ndltd.org/standards/metadata/etdms/1.1/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.ndltd.org/standards/metadata/etdms/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdms11.xsd http://purl.org/dc/elements/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdmsdc.xsd">
          <dc:contributor>Ahuja, Narendra</dc:contributor>
          <dc:contributor>Ahuja, Narendra</dc:contributor>
          <dc:contributor>Huang, Thomas S.</dc:contributor>
          <dc:contributor>Forsyth, David A.</dc:contributor>
          <dc:contributor>Hasegawa-Johnson, Mark A.</dc:contributor>
          <dc:contributor>Hoiem, Derek W.</dc:contributor>
          <dc:creator>Akbas, Emre</dc:creator>
          <dc:date>2011-08-26T15:22:42Z</dc:date>
          <dc:date>2011-08-26T15:22:42Z</dc:date>
          <dc:date>2013-08-27T10:00:22Z</dc:date>
          <dc:date>2011-08-26T15:22:42Z</dc:date>
          <dc:date>2011-08</dc:date>
          <dc:description>This dissertation is about extracting as well as making use of  the structure and hierarchy present in images. We develop a new low-level, multiscale, hierarchical image segmentation algorithm designed to detect image regions regardless of their shapes, sizes, and levels of interior homogeneity.  We model a region as a connected set of pixels that is surrounded by ramp edge discontinuities where the magnitude of these discontinuities is large compared to the variation inside the region. Each region is associated with a scale depending on the magnitude of the weakest part of its boundary. Traversing through the range of all possible scales, we obtain all regions
present in the image. Regions strictly merge as the scale increases; hence a
tree is formed where the root node corresponds to the whole image, and nodes close to the root along a path are large, while their children nodes are smaller and
capture embedded details.
To evaluate the accuracy and precision of our algorithm, as well as to compare
it to the existing algorithms, we develop a new benchmark dataset for low-level image segmentation. In this benchmark, small patches of many images are hand-segmented by human subjects. We provide evaluation methods
for both boundary-based and region-based performance of algorithms.  We show that our proposed algorithm performs better than the
existing low-level segmentation algorithms on this benchmark. 
Next, we investigate the segmentation-based statistics of natural images. Such
statistics capture geometric and topological properties of images, which is not
possible to obtain using pixel-, patch-, or subband-based methods. We compile
and use segmentation statistics from a large number of images, and propose a
Markov random field based  model for estimating them. Our estimates confirm some
of the previous statistical properties of natural images as well as yield new
ones.  To demonstrate the value of the statistics, we successfully use them as
priors in image classification and semantic image segmentation.
We also investigate the importance of different visual cues to
describe image regions for solving the region correspondence problem. We design
and develop
psychophysical experiments to learn the weights of different cues by evaluating
their impact on binocular fusibility by human subjects. Using a head-mounted
display, we show a set of elliptical regions to one eye and slightly different
versions of the same set of regions to the other eye of human subjects. We then
ask them whether the ellipses fuse or not. By systematically varying the
parameters of the elliptical shapes, and testing for fusion,  we learn a
perceptual distance function between two elliptical regions. We evaluate this
function on ground-truth stereo image pairs. 
Finally, we propose a novel multiple instance learning (MIL) method. In MIL,
in contrast to classical supervised learning, the entities to be classified
are called bags, each of which contains an arbitrary number of elements called
instances. 
We propose an additive model for bag classification where we 
exploit the idea of searching for discriminative instances, which we call
prototypes.  We show that our bag-classifier can be learned in a boosting
framework, leading to an iterative algorithm, which learns prototype-based
weak learners that are linearly combined. At each iteration of our proposed
method, we search for a new prototype so as to maximally discriminate between
the positive and negative bags, which are themselves weighted according to how
well they were discriminated in earlier iterations. Unlike previous instance selection based MIL
methods, we do not restrict the prototypes to a discrete set of training
instances but allow them to take arbitrary values in the instance feature space.
We also do not restrict the total number of prototypes and the number of
selected-instances per bag; these quantities are completely data-driven. We show
that our method outperforms state-of-the-art MIL methods on a number of benchmark
datasets. We also apply our method to large-scale image classification,
where we show that the automatically selected prototypes map to visually meaningful image regions.</dc:description>
          <dc:description>Item withdrawn by Mark Zulauf (zulauf@illinois.edu) on 2011-07-12T19:05:22Z
Item was in collections:
University of Illinois Theses &amp; Dissertations (ID: 1)
No. of bitstreams: 1
Akbas_Emre.pdf: 13019843 bytes, checksum: 07786039448ea8ce8c9389ca8f6ba0a8 (MD5)</dc:description>
          <dc:description>Made available in DSpace on 2011-08-26T15:22:42Z (GMT). No. of bitstreams: 2
Akbas_Emre.pdf: 13019843 bytes, checksum: 07786039448ea8ce8c9389ca8f6ba0a8 (MD5)
license.txt: 4059 bytes, checksum: 20ce3dd81102f792773aed2db3c84cd7 (MD5)</dc:description>
          <dc:description>Item marked as restricted to the 'UIUC Users [automated]' Group (id=2) by William Ingram (wingram2@illinois.edu) on 2011-08-26T15:25:57Z
Item is restricted until 2013-08-26T15:25:28Z</dc:description>
          <dc:description>Item reinstated by Sarah Shreeves (sshreeve@illinois.edu) on 2013-08-27T10:00:22Z
Item was in collections:
University of Illinois Dissertations and Theses (ID: 204)
Dissertations and Theses - Electrical and Computer Engineering (ID: 446)
No. of bitstreams: 3
Akbas_Emre.pdf.txt: 232682 bytes, checksum: 167459abc7e87e6ee93ec7953d6252ea (MD5)
Akbas_Emre.pdf: 13019843 bytes, checksum: 07786039448ea8ce8c9389ca8f6ba0a8 (MD5)
license.txt: 4059 bytes, checksum: 20ce3dd81102f792773aed2db3c84cd7 (MD5)</dc:description>
          <dc:description>Item released from any restrictions by Sarah Shreeves (sshreeve@illinois.edu) on 2013-08-27T10:00:22Z</dc:description>
          <dc:identifier>http://hdl.handle.net/2142/26317</dc:identifier>
          <dc:language>en</dc:language>
          <dc:rights>Copyright 2011 Emre Akbas</dc:rights>
          <dc:subject>computer vision</dc:subject>
          <dc:subject>image processing</dc:subject>
          <dc:subject>image segmentation</dc:subject>
          <dc:subject>machine learning</dc:subject>
          <dc:subject>segmentation benchmark</dc:subject>
          <dc:subject>natural image statistics</dc:subject>
          <dc:subject>image classification</dc:subject>
          <dc:subject>scene classification</dc:subject>
          <dc:subject>binocular fusion</dc:subject>
          <dc:subject>region matching</dc:subject>
          <dc:subject>multiple instance learning (mil)</dc:subject>
          <dc:subject>mis-boost</dc:subject>
          <dc:subject>pascal voc</dc:subject>
          <dc:subject>Pattern Analysis Statistical Modeling and Computational Learning (PASCAL)</dc:subject>
          <dc:subject>Visual Object Classes (VOC)</dc:subject>
          <dc:title>Generation and analysis of segmentation trees for natural images</dc:title>
          <degree>
            <department>Electrical &amp; Computer Eng</department>
            <departmentCode>1933</departmentCode>
            <discipline>Electrical &amp; Computer Engr</discipline>
            <disciplineCode>1200</disciplineCode>
            <grantor>University of Illinois at Urbana-Champaign</grantor>
            <level>Dissertation</level>
            <name>Ph.D.</name>
            <program>PHD:Electr &amp; Computer Eng-UIUC</program>
            <programCode>10KS1200PHD</programCode>
          </degree>
        </thesis>
      </metadata>
    </record>
  </GetRecord>
</OAI-PMH>
