<?xml version="1.0" encoding="UTF-8"?><?xml-stylesheet type="text/xsl" href="static/style.xsl"?><OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd"><responseDate>2026-09-18T20:55:41Z</responseDate><request verb="GetRecord" identifier="oai:dspace.mit.edu:1721.1/164851" metadataPrefix="dim">https://dspace.mit.edu/server/oai/request</request><GetRecord><record><header><identifier>oai:dspace.mit.edu:1721.1/164851</identifier><datestamp>2026-02-13T03:49:22Z</datestamp><setSpec>com_1721.1_7582</setSpec><setSpec>com_1721.1_7581</setSpec><setSpec>col_1721.1_131023</setSpec></header><metadata><dim:dim xmlns:dim="http://www.dspace.org/xmlns/dspace/dim" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:doc="http://www.lyncode.com/xoai" xsi:schemaLocation="http://www.dspace.org/xmlns/dspace/dim http://www.dspace.org/schema/dim.xsd">
   <dim:field mdschema="dc" element="contributor" qualifier="advisor">Ghassemi, Marzyeh</dim:field>
   <dim:field mdschema="dc" element="contributor" qualifier="author">Li, Angela</dim:field>
   <dim:field mdschema="dc" element="contributor" qualifier="department">Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="accessioned">2026-02-12T17:14:32Z</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="available">2026-02-12T17:14:32Z</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="issued">2025-09</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="submitted">2025-09-15T14:56:32.588Z</dim:field>
   <dim:field mdschema="dc" element="identifier" qualifier="uri">https://hdl.handle.net/1721.1/164851</dim:field>
   <dim:field mdschema="dc" element="description" qualifier="abstract">Large Language Models (LLMs) have demonstrated remarkable success in many applications; however, their reliability remains a critical concern, especially in high-stakes domains such as healthcare, finance, and law. Uncertainty Quantification (UQ) is essential for assessing LLM outputs and ensuring trust. However, existing UQ methods for LLMs face challenges: high computational costs, difficulties in handling unstructured outputs, and limited generalizability. This thesis addresses these challenges by proposing a systematic investigation into robust and efficient UQ methodologies tailored for LLMs. Specifically, this work focuses on: (1) analyzing probing methods to determine how hidden layers encode information relevant to uncertainty and accuracy, (2) developing novel UQ metrics that strongly correlate with actual model performance, and (3) designing computationally efficient pipelines to make UQ practical for real-world applications. By bridging these gaps, this research aims to establish UQ as a reliable tool for evaluating and improving the trustworthiness of LLM outputs, facilitating their safe and effective deployment in critical domains.</dim:field>
   <dim:field mdschema="dc" element="description" qualifier="degree">MNG</dim:field>
   <dim:field mdschema="dc" element="publisher">Massachusetts Institute of Technology</dim:field>
   <dim:field mdschema="dc" element="rights">In Copyright - Educational Use Permitted</dim:field>
   <dim:field mdschema="dc" element="rights">Copyright retained by author(s)</dim:field>
   <dim:field mdschema="dc" element="rights" qualifier="uri">https://rightsstatements.org/page/InC-EDU/1.0/</dim:field>
   <dim:field mdschema="dc" element="title">Efficient Uncertainty Quantification of Large Language&#xd;
Models</dim:field>
   <dim:field mdschema="dc" element="type">Thesis</dim:field>
   <dim:field mdschema="dc" element="format" qualifier="mimetype">application/pdf</dim:field>
   <dim:field mdschema="mit" element="thesis" qualifier="degree">Master</dim:field>
   <dim:field mdschema="thesis" element="degree" qualifier="name" />
   <dim:field mdschema="dspace" element="entity" qualifier="type">Publication</dim:field>
   <dim:field mdschema="others" element="access-status">unknown</dim:field>
   <dim:field mdschema="others" element="access-status">unknown</dim:field>
   <dim:field mdschema="cerif" element="openaire" authority="" confidence="-1">&lt;Publication xmlns="https://www.openaire.eu/cerif-profile/1.1/" id="90203206-ee4f-46ae-aa6e-6fb55ffb6a34">
	&lt;Type xmlns="https://www.openaire.eu/cerif-profile/vocab/COAR_Publication_Types">http://purl.org/coar/resource_type/c_1843&lt;/Type>
   	&lt;Title>Efficient Uncertainty Quantification of Large Language&#xd;
Models&lt;/Title>
   	&lt;PublishedIn>
    	&lt;Publication>
      	&lt;/Publication>
   	&lt;/PublishedIn>
   	&lt;PublicationDate>2025-09&lt;/PublicationDate>
   	&lt;Authors>
      	&lt;Author>
        	&lt;DisplayName>Li, Angela&lt;/DisplayName>
         	&lt;Affiliation>
         		&lt;OrgUnit>
         		&lt;/OrgUnit>
         	&lt;/Affiliation>
      	&lt;/Author>
	&lt;/Authors>
   	&lt;Editors>
	&lt;/Editors>
    &lt;Publishers>
        &lt;Publisher>
            &lt;DisplayName>Massachusetts Institute of Technology&lt;/DisplayName>
            &lt;OrgUnit />
        &lt;/Publisher>
    &lt;/Publishers>
    &lt;License>https://rightsstatements.org/page/InC-EDU/1.0/&lt;/License>
   	&lt;Abstract>Large Language Models (LLMs) have demonstrated remarkable success in many applications; however, their reliability remains a critical concern, especially in high-stakes domains such as healthcare, finance, and law. Uncertainty Quantification (UQ) is essential for assessing LLM outputs and ensuring trust. However, existing UQ methods for LLMs face challenges: high computational costs, difficulties in handling unstructured outputs, and limited generalizability. This thesis addresses these challenges by proposing a systematic investigation into robust and efficient UQ methodologies tailored for LLMs. Specifically, this work focuses on: (1) analyzing probing methods to determine how hidden layers encode information relevant to uncertainty and accuracy, (2) developing novel UQ metrics that strongly correlate with actual model performance, and (3) designing computationally efficient pipelines to make UQ practical for real-world applications. By bridging these gaps, this research aims to establish UQ as a reliable tool for evaluating and improving the trustworthiness of LLM outputs, facilitating their safe and effective deployment in critical domains.&lt;/Abstract>
	&lt;Access xmlns="http://purl.org/coar/access_right" 
    >
    &lt;/Access>
&lt;/Publication>
</dim:field>
</dim:dim>
</metadata></record></GetRecord></OAI-PMH>