Project Grant 2339746

Award Date 5/1/24
Completion Date 4/30/29
Dollars Obligated $246K
Federal Grant Program
47.070
Assistance Type
Project Grant
Place of Performance
College Park, MD 20742, USA
Similar Awards
This Project Grant award from the National Science Foundation (NSF) Computer and Information Science and Engineering (CISE) program (CFDA 47.070) provides $800,000 to the University of Illinois Urbana-Champaign to enhance the safety of large language models (LLMs) used in high-stakes applications. The technical aims of the project include: (1) developing "robust-confidence safety" to ensure LLMs appropriately respond to out-of-distribution scenarios and rare events, (2) enforcing...
This $100,000 Project Grant award from the National Science Foundation (NSF) under the Computer and Information Science and Engineering (CISE) program (CFDA 47.070) will support a planning effort for a proposed large-scale, impactful CISE core project. The grant will allow the University of Louisville research team to investigate novel algorithms for incorporating large language models into recommendation systems, with a focus on addressing fairness, robustness, and trustworthiness. The planning...
This $175,000 two-year Project Grant from the National Science Foundation's Division of Information and Intelligent Systems, under the Computer and Information Science and Engineering program (CFDA 47.070), will fund research at the University of Southern California to develop more explainable artificial intelligence agents with common sense capabilities. The grant aims to advance techniques for AI agents to reason in novel situations, complete open-world narratives, and provide commonsense...
This $300,000 Project Grant awarded by the National Science Foundation (NSF) under the NSF Technology, Innovation, and Partnerships (CFDA 47.084) program supports the development of an evaluation methodology to measure the impacts of applying large language model (LLM)-based tools within federal, state, and local government programs. The goal is to assess whether LLMs can help reduce complexity and potential for human error in government service delivery by assisting staff with tasks like data...
This $174,995 Project Grant awarded by the National Science Foundation (NSF) under the Computer and Information Science and Engineering (CFDA 47.070) program supports research to understand and optimize the role of human intelligence in data integration and discovery pipelines, especially in the context of emerging large language models (LLMs) like ChatGPT. The project will investigate fundamental questions about human involvement in these data processes, uncover relevant human biases, and...
This $325,925 federal Project Grant award from the National Science Foundation's (NSF) Computer and Information Science and Engineering (CISE) program (CFDA 47.070) supports the development of a computational framework for addressing data challenges related to fairness in machine learning. The project aims to integrate "model-centric" and "data-centric" approaches to enhance the generalizability, trustworthiness, and fairness of machine learning algorithms, particularly in...
This National Science Foundation (NSF) project grant, awarded under the Computer and Information Science and Engineering (CISE) program (CFDA 47.070), supports research to develop information-theoretic approaches for ensuring fairness and explainability in high-stakes machine learning applications. The $125,886 award to the University of Maryland, College Park aims to advance the foundations of ethical and socially-responsible machine learning. The project will leverage information theory...
This National Science Foundation (NSF) Computer and Information Science and Engineering (CISE) Federal Grant Program (CFDA 47.070) award of $799,999 to the University of Washington provides funding for a 4-year collaborative research project focused on developing tools and methods to leverage large language models (LLMs) in support of divergent, convergent, and cooperative work. The key objectives of the project are to: 1) Develop design guidance and workflows for building more reliable and...
This Project Grant award of $790,797 from the National Science Foundation (NSF) Computer and Information Science and Engineering (CISE) program (CFDA 47.070) supports research at the University of California, San Diego (UCSD) to investigate "Reasoning About Multiplicity in the Machine Learning Pipeline." The project aims to develop formal techniques and frameworks to understand how multiplicity (the existence of multiple equally good models) in the machine learning pipeline,...
This $625,000 National Science Foundation project grant supports research at the University of Maryland, College Park toward developing fair machine learning algorithms and applications. Specifically, the grant funds the "Toward Fair Decision Making and Resource Allocation with Application to AI-Assisted Graduate Admission and Degree Completion" project. The research aims to design artificial intelligence systems that make admission and resource allocation decisions for graduate...

CAREER: ROBUST, FAIR, AND CULTURALLY AWARE COMMONSENSE REASONING IN NATURAL LANGUAGE -RECENT ADVANCES IN ARTIFICIAL INTELLIGENCE HAVE LED TO THE PROLIFERATION OF LARGE LANGUAGE MODELS (LLMS). LLMS ARE MODELS THAT CANE BE USED FOR INTERACTIONS WITH HUMAN USERS THROUGH WRITTEN LANGUAGE; FOR EXAMPLE, A USER INPUTS AN INSTRUCTION OR QUESTION IN ENGLISH TO THE LLM-BASED PROGRAM, AND THE LLM OUTPUTS A RESPONSE IN FLUENT ENGLISH. WITH THESE LINGUISTIC CAPABILITIES, LLMS ARE BEING DEVELOPED FOR USE IN APPLICATIONS THAT ARE BOTH UBIQUITOUS (E.G., INTERNET SEARCH, CUSTOMER SUPPORT, WRITING TOOLS) AND HIGH-STAKES (E.G., MENTAL HEALTH CARE, CLASSROOM EDUCATION, ASSISTIVE TECHNOLOGY FOR PEOPLE WITH DISABILITIES). DESPITE THEIR GROWING ADOPTION, MANY FUNDAMENTAL PROPERTIES OF LLMS AREN?T YET WELL UNDERSTOOD, AND PRESSING QUESTIONS REMAIN ABOUT WHEN AND WHETHER LLMS CAN BE ENTRUSTED WITH SUCH IMPORTANT TASKS. FOR EXAMPLE, WHEN INSTRUCTED TO MAKE SIMPLE PREDICTIONS ABOUT EVERY-DAY SITUATIONS, LIKE COOKING A MEAL OR RIDING IN A VEHICLE, LLMS CAN MAKE STRANGE AND SURPRISING ERRORS, EXHIBITING CONCERNING LAPSES IN BASIC COMMON SENSE JUDGMENT AND REASONING ABILITIES. ADDITIONALLY, THESE PREDICTIONS MADE BY LLMS CAN REFLECT SOCIAL STEREOTYPES AND CULTURAL ASSUMPTIONS WHICH, AT BEST, LIMIT THE USEFULNESS OF THE TECHNOLOGY FOR CERTAIN POPULATIONS AND, AT WORST, CAUSE ACTIVE HARM. THIS PROJECT SEEKS TO ADDRESS UNFAIRNESS AND BIAS DUE TO STEREOTYPING AND CULTURAL CONTEXT BY PROPOSING A GENERALIZED FRAMEWORK FOR DEFEASIBLE COMMONSENSE INFERENCE IN NATURAL LANGUAGE IN WHICH A SYSTEM COMPARES TWO SIMILAR SITUATIONS WITH RESPECT TO THEIR SUPPORT FOR A GIVEN INFERENCE. THE PROPOSED WORK AIMS AT DEVELOPING SCIENTIFIC METHODS TO MEASURE AND IMPROVE THE ABILITIES OF LLMS TO (1) REASON CORRECTLY ABOUT EVERY-DAY SITUATIONS, (2) DO SO IN A MANNER THAT IS FAIR AND UNPREJUDICED, AND (3) ADAPT THESE REASONING ABILITIES ACROSS SPECIFIC CULTURAL CONTEXTS. BY MEASURING THESE FUNDAMENTAL CAPABILITIES OF LLMS, WE CAN BETTER UNDERSTAND AND MITIGATE THE RISKS OF APPLYING THIS TECHNOLOGY IN HIGH-STAKES SETTINGS. THE THREE PHASES OF THE PROJECT FOCUS ON THE (1) ROBUSTNESS, (2) SOCIAL FAIRNESS, AND (3) CULTURAL AWARENESS DIMENSIONS OF REASONING IN LLMS. THE PROJECT ASSUMES A BASIC TASK FORMULATION IN WHICH A SITUATION DESCRIPTION IS PROVIDED TO AN LLM (E.G., ?SOMEONE DROPS A GLASS?), AND THE LLM MUST EITHER EVALUATE A POSSIBLE INFERENCE, OR GENERATE AN INFERENCE FROM SCRATCH (?THE GLASS BREAKS?). IN PHASE 1, METHODS WILL BE DEVELOPED TO AUTOMATICALLY MANIPULATE SITUATION DESCRIPTIONS IN ORDER TO TRAIN AND EVALUATE AN LLM?S ABILITY TO MAKE NUANCED INFERENCES, WITH THE GOAL OF LEARNING TO DISTINGUISH WHICH FACTORS INFLUENCE A PARTICULAR INFERENCE AND WHICH ONES DO NOT (E.G., WHEN TRYING TO PREDICT IF A DROPPED GLASS IS GOING TO BREAK, THE THICKNESS OF THE GLASS MATTERS BUT THE COLOR OF THE GLASS DOES NOT.) IN PHASE 2, METHODS WILL BE DEVELOPED TO AUTOMATICALLY TEST WHETHER LLMS MAKE SOCIALLY FAIR INFERENCES, FOR EXAMPLE VIA NAME SUBSTITUTION TESTS, AND TO INTERVENE WHEN A PROPOSED OUTPUT IS DETECTED AS UNFAIR. IN PHASE 3, SURVEY PARTICIPANTS FROM THE U.S. AND GHANA WILL ANSWER MULTIPLE STAGES OF QUESTIONS ABOUT EVERY-DAY SITUATIONS; THE COLLECTED DATA WILL BE USED TO DEVELOP EVALUATION QUESTIONS FOR A CASE STUDY ON THE ADAPTABILITY OF LLMS ACROSS THESE TWO CULTURAL SETTINGS. FOR EACH PHASE OF THE PROJECT, THE RESULTING DATASETS, METHODS, AND SCIENTIFIC FINDINGS WILL BE MADE AVAILABLE TO THE PUBLIC. THIS AWARD REFLECTS NSF'S STATUTORY MISSION AND HAS BEEN DEEMED WORTHY OF SUPPORT THROUGH EVALUATION USING THE FOUNDATION'S INTELLECTUAL MERIT AND BROADER IMPACTS REVIEW CRITERIA.- SUBAWARDS ARE NOT PLANNED FOR THIS AWARD.

Posted 4/9/24, 12:00 AM