-
2024How can we precisely estimate a large language model’s (LLM) accuracy on questions belonging to a specific topic within a larger question-answering dataset? The standard direct estimator, which averages the model’s accuracy on the questions in each subgroup, may exhibit high variance for subgroups (topics) with small sample sizes. Synthetic regression modeling, which leverages the model’s accuracy on questions
-
2024Modern language models (LMs) need to follow human instructions while being faithful; yet, they often fail to achieve both. Here, we provide concrete evidence of a trade-off between instruction following (i.e., follow open-ended instructions) and faithfulness (i.e., ground responses in given context) when training LMs with these objectives. For instance, fine-tuning LLaMA-7B on instruction following datasets
-
2024Learning of preference models from human feedback has been central to recent advances in artificial intelligence. Motivated by the cost of obtaining high-quality human annotations, we study efficient human preference elicitation for learning preference models. The key idea in our work is to generalize optimal designs, a methodology for computing optimal information-gathering policies, to questions with
-
2024Large language model advancements have enabled the development of multi-agent frameworks to tackle complex, real-world problems such as to automate tasks that require interactions with diverse tools, reasoning, and human collaboration. We present MARCO, a Multi-Agent Real-time Chat Orchestration framework for automating tasks using LLMs. MARCO addresses key challenges in utilizing LLMs for complex, multi-step
-
2024While the Transformer architecture has achieved remarkable success across various domains, a thorough theoretical foundation explaining its optimization dynamics is yet to be fully developed. In this study, we aim to bridge this understanding gap by answering the following two core questions: (1) Which types of Transformer architectures allow Gradient Descent (GD) to achieve guaranteed convergence? and
Related content
-
December 22, 2022Google JAX Python library implementation and new topics added; volume 1 of book to be published by Cambridge University Press.
-
December 15, 2022New, free offering provides students of any level practical skills and code examples for every stage, from the machine learning problem all the way to deployment.
-
December 14, 2022Alexa Fund company’s assisted reality tech could unlock speech for hundreds of millions of people who struggle to communicate.
-
December 12, 2022Vice president of ML and AI Services says more than 100,000 customers are doing machine learning on AWS.
-
December 12, 2022Vice president Bratin Saha reflects on the past and future of Amazon Web Services’ machine learning tools and AI services.
-
December 07, 2022Learn about a real-time continual, lifelong learning system that trains machine learning models using production data at scale.