<?xml version="1.0" encoding="UTF-8"?>
<?xml-stylesheet type="text/xsl" href="/oai-pmh.xsl"?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-09-19T01:05:49Z</responseDate>
  <request identifier="oai:www.ideals.illinois.edu:2142/110569" metadataPrefix="etdms" verb="GetRecord">https://www.ideals.illinois.edu/oai-pmh</request>
  <GetRecord>
    <record>
      <header>
        <identifier>oai:www.ideals.illinois.edu:2142/110569</identifier>
        <datestamp>2023-07-11</datestamp>
        <setSpec>col_2142_5131</setSpec>
        <setSpec>col_2142_10761</setSpec>
        <setSpec>com_2142_5130</setSpec>
        <setSpec>com_2142_10755</setSpec>
        <setSpec>com_2142_234</setSpec>
      </header>
      <metadata>
        <thesis xmlns="http://www.ndltd.org/standards/metadata/etdms/1.1/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.ndltd.org/standards/metadata/etdms/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdms11.xsd http://purl.org/dc/elements/1.1/ http://www.ndltd.org/standards/metadata/etdms/1.1/etdmsdc.xsd">
          <dc:contributor>Zhai, Chengxiang</dc:contributor>
          <dc:creator>Cheng, Xiang</dc:creator>
          <dc:date>2021-09-17T01:11:17Z</dc:date>
          <dc:date>2021-09-17T01:11:17Z</dc:date>
          <dc:date>2021-04-27</dc:date>
          <dc:date>2021-05</dc:date>
          <dc:description>Structure prediction (SP) tasks are important in natural language understanding in the sense that they provide complex and structured knowledge of the text. Recently, some unified text-to-text transformer models like T5 and TANL have produced competitive results on SP tasks. These models convert SP tasks into a seq2seq problem, where a transformer is used to generate sequences with special tokens representing the extracted spans, labels, and relationships. Compared to many popular Natural Language Understanding models that are designed specifically for the task, the output of the text-to-text transformer is more flexible.
With proper format, it could be trained on multiple tasks together and take advantage of the shared knowledge between tasks. To better understand how these models achieve better performance by multi-task learning, we designed several experiments to measure the knowledge transfer ability of a recently proposed model, TANL. In our experiments, we found that the multi-head attention in the decoder can capture the relationship between tasks which leads to performance improvement. Another finding is that TANL may produce many outputs with invalid format when trained from scratch, and starting from a T5 pre-trained model helps to mitigate this problem. 
Based on these observations and some new intuitions, we proposed an improved version of TANL called SDCT5 (step decomposed and constrained text-to-text Transformer). Preliminary experiment results show that our model can achieve better performance on SP tasks compared to TANL and benefit more from multi-task learning.</dc:description>
          <dc:description>Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2021-09-16 without embargo terms</dc:description>
          <dc:description>The student, Xiang Cheng, accepted the attached license on 2021-04-24 at 10:19.</dc:description>
          <dc:description>The student, Xiang Cheng, submitted this Thesis for approval on 2021-04-24 at 10:28.</dc:description>
          <dc:description>This Thesis was approved for publication on 2021-04-27 at 09:34.</dc:description>
          <dc:description>DSpace SAF Submission Ingestion Package generated from Vireo submission #16544 on 2021-09-16 at 16:47:31</dc:description>
          <dc:description>Made available in DSpace on 2021-09-17T01:11:17Z (GMT). No. of bitstreams: 2
CHENG-THESIS-2021.pdf: 1242117 bytes, checksum: c62d59918b73c76eb5fc07f0b10311e2 (MD5)
LICENSE.txt: 4208 bytes, checksum: d01d4fe5f48c513e9a750014751422ec (MD5)
  Previous issue date: 2021-04-27</dc:description>
          <dc:format>application/pdf</dc:format>
          <dc:identifier>http://hdl.handle.net/2142/110569</dc:identifier>
          <dc:language>en</dc:language>
          <dc:rights>Copyright 2021 Xiang Cheng</dc:rights>
          <dc:subject>natural language processing</dc:subject>
          <dc:subject>structure prediction</dc:subject>
          <dc:subject>multi-task learning</dc:subject>
          <dc:title>A deeper look into multi-task learning ability of unified text-to-text transformer</dc:title>
          <dc:type>text</dc:type>
          <dc:type>Thesis</dc:type>
          <degree>
            <department>Computer Science</department>
            <discipline>Computer Science</discipline>
            <grantor>University of Illinois at Urbana-Champaign</grantor>
            <level>Thesis</level>
            <name>M.S.</name>
          </degree>
        </thesis>
      </metadata>
    </record>
  </GetRecord>
</OAI-PMH>
