<?xml version="1.0" encoding="UTF-8"?>
<?xml-stylesheet type="text/xsl" href="/oai-pmh.xsl"?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-09-20T05:26:16Z</responseDate>
  <request identifier="oai:www.ideals.illinois.edu:2142/92750" metadataPrefix="etdms" verb="GetRecord">https://www.ideals.illinois.edu/oai-pmh</request>
  <GetRecord>
    <record>
      <header>
        <identifier>oai:www.ideals.illinois.edu:2142/92750</identifier>
        <datestamp>2023-07-11</datestamp>
        <setSpec>col_2142_8888</setSpec>
        <setSpec>col_2142_5131</setSpec>
        <setSpec>com_2142_8887</setSpec>
        <setSpec>com_2142_234</setSpec>
        <setSpec>com_2142_5130</setSpec>
      </header>
      <metadata>
        <thesis xmlns="http://www.ndltd.org/standards/metadata/etdms/1.1/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.ndltd.org/standards/metadata/etdms/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdms11.xsd http://purl.org/dc/elements/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdmsdc.xsd">
          <dc:contributor>Ahuja, Narendra</dc:contributor>
          <dc:contributor>Ahuja, Narendra</dc:contributor>
          <dc:contributor>Huang, Thomas S.</dc:contributor>
          <dc:contributor>Do, Minh N.</dc:contributor>
          <dc:contributor>Hasegawa-Johnson, Mark</dc:contributor>
          <dc:contributor>Hoiem, Derek</dc:contributor>
          <dc:creator>Huang, Jia-Bin</dc:creator>
          <dc:date>2016-11-10T17:50:06Z</dc:date>
          <dc:date>2016-11-10T17:50:06Z</dc:date>
          <dc:date>2016-07-06</dc:date>
          <dc:date>2016-08</dc:date>
          <dc:description>"The past decade has witnessed remarkable progress in image-based, data-driven vision and graphics. However, existing approaches often treat the images as pure 2D signals and not as a 2D projection of the physical 3D world. As a result, a lot of training examples are required to cover sufficiently diverse appearances and inevitably suffer from limited generalization capability. In this thesis, I propose ""inference-by-composition"" approaches to overcome these limitations by modeling and interpreting visual signals in terms of physical surface, object, and scene. I show how we can incorporate physically grounded constraints such as scene-specific geometry in a non-parametric optimization framework for (1) revealing the missing parts of an image due to removal of a foreground or background element, (2) recovering high spatial frequency details that are not resolvable in low-resolution observations. I then extend the framework from 2D images to handle spatio-temporal visual data (videos). I demonstrate that we can convincingly fill spatio-temporal holes in a temporally coherent fashion by jointly reconstructing the appearance and motion. Compared to existing approaches, our technique can synthesize physically plausible contents even in challenging videos. For visual analysis, I apply stereo camera constraints for discovering multiple approximately linear structures in extremely noisy videos with an ecological application to bird migration monitoring at night. The resulting algorithms are simple and intuitive while achieving state-of-the-art performance without the need of training on an exhaustive set of visual examples."</dc:description>
          <dc:description>Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2016-11-09 without embargo terms</dc:description>
          <dc:description>The student, Jia-Bin Huang, accepted the attached license on 2016-07-04 at 02:00.</dc:description>
          <dc:description>The student, Jia-Bin Huang, submitted this Dissertation for approval on 2016-07-04 at 02:17.</dc:description>
          <dc:description>This Dissertation was approved for publication on 2016-07-06 at 14:09.</dc:description>
          <dc:description>DSpace SAF Submission Ingestion Package generated from Vireo submission #9749 on 2016-11-09 at 10:22:39</dc:description>
          <dc:description>Made available in DSpace on 2016-11-10T17:50:06Z (GMT). No. of bitstreams: 3
HUANG-DISSERTATION-2016.pdf: 41822301 bytes, checksum: 792c5fe8a31458896c5c2e37d284e871 (MD5)
LICENSE.txt: 4210 bytes, checksum: 6d24520218c46f6c63a5d4206401a794 (MD5)
PROQUEST_LICENSE.txt: 4556 bytes, checksum: 7476bec9b87a706036712e765c2b7fd7 (MD5)
  Previous issue date: 2016-07-06</dc:description>
          <dc:format>application/pdf</dc:format>
          <dc:identifier>http://hdl.handle.net/2142/92750</dc:identifier>
          <dc:language>en</dc:language>
          <dc:rights>Copyright 2016 Jia-Bin Huang</dc:rights>
          <dc:subject>Computer vision</dc:subject>
          <dc:subject>Visual synthesis</dc:subject>
          <dc:subject>patch-based optimization</dc:subject>
          <dc:subject>image completion</dc:subject>
          <dc:subject>image super-resolution</dc:subject>
          <dc:subject>video completion</dc:subject>
          <dc:subject>visual tracking</dc:subject>
          <dc:title>Visual analysis and synthesis with physically grounded constraints</dc:title>
          <dc:type>text</dc:type>
          <dc:type>text</dc:type>
          <degree>
            <department>Electrical &amp; Computer Eng</department>
            <discipline>Electrical &amp; Computer Engr</discipline>
            <grantor>University of Illinois at Urbana-Champaign</grantor>
            <level>Dissertation</level>
            <name>Ph.D.</name>
          </degree>
        </thesis>
      </metadata>
    </record>
  </GetRecord>
</OAI-PMH>
