31% of US adults use generative AI for healthcare π€―But most AI systems answer questions assertivelyβeven when they donβt have the necessary context. Introducing #MediQ a framework that enables LLMs to recognize uncertaintyπ€and ask the right questionsβwhen info is missing: π§΅
06.12.2024 22:51 β
π 68
π 14
π¬ 2
π 2
Excited to be at #ICLR2025 π€©
I'll be giving an oral presentation for Creativity Index on Fri 25th 11:06, Garnet 212&219 ποΈ
I'll also be presenting posters:
πExploreToM, Sat 26th 10:00, Hall 3 + 2B #49
πCreativityIndex, Fri 25th 15:00, Hall 3 + 2B #618
Hope to see you there!
24.04.2025 02:25 β
π 8
π 1
π¬ 0
π 0
A screenshot of the first page of the paper, containing the paper title: Finding Flawed Fictions: Evaluating Complex Reasoning in Language Models via Plot Hole Detection and the names of the authors: Kabir Ahuja, Melanie Sclar, and Yulia Tsvetkov. All the three authors are from CSE department in the University of Washington in Seattle, USA. They can be reached at {kahuja,msclar,yuliats}@cs.washington.edu
π’ New Paper!
Tired π΄ of reasoning benchmarks full of math & code? In our work we consider the problem of reasoning for plot holes in stories -- inconsistencies in a storyline that break the internal logic or rules of a storyβs world π
W @melaniesclar.bsky.social, and @tsvetshop.bsky.social
1/n
22.04.2025 18:50 β
π 10
π 4
π¬ 1
π 1
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
High-quality training data has proven crucial for developing performant large language models (LLMs). However, commercial LLM providers disclose few, if any, details about the data used for training. ...
Want to know what training data has been memorized by models like GPT-4?
We propose information-guided probes, a method to uncover memorization evidence in *completely black-box* models,
without requiring access to
π
ββοΈ Model weights
π
ββοΈ Training data
π
ββοΈ Token probabilities π§΅ (1/5)
21.03.2025 19:08 β
π 97
π 27
π¬ 4
π 8
π¨New Paper! So o3-mini and R1 seem to excel on math & coding. But how good are they on other domains where verifiable rewards are not easily available, such as theory of mind (ToM)? Do they show similar behavioral patterns? π€ What if I told you it's...interesting, like the below?π§΅
20.02.2025 17:34 β
π 22
π 5
π¬ 3
π 1
We are launching HALoGENπ‘, a way to systematically study *when* and *why* LLMs still hallucinate.
New work w/ Shrusti Ghela*, David Wadden, and Yejin Choi π«
π Paper: arxiv.org/abs/2501.08292
π Code/Data: github.com/AbhilashaRav...
π Website: halogen-hallucinations.github.io π§΅ [1/n]
31.01.2025 18:27 β
π 34
π 8
π¬ 3
π 4
Iβm on the academic job market this year! Iβm completing my @uwcse.bsky.social @uwnlp.bsky.social Ph.D. (2025), focusing on overcoming LLM limitations like hallucinations, by building new LMs.
My Ph.D. work focuses on Retrieval-Augmented LMs to create more reliable AI systems π§΅
04.12.2024 13:26 β
π 71
π 17
π¬ 3
π 2
Want to predict the task performance of LMs before pretraining them?
We develop task scaling laws and model ladders, which predict the accuracy on individual tasks by OLMo 2 7B & 13B models within 2 points of absolute error. The cost is 1% of the compute used to pretrain them.
09.12.2024 17:07 β
π 33
π 14
π¬ 2
π 0
poster for paper
excited to be at #NeurIPS2024! I'll be presenting our data mixture inference attack ποΈ Thu 4:30pm w/ @jon.jon.ke β stop by to learn what trained tokenizers reveal about LLM development (βΌοΈ) and chat about all things tokenizers.
π arxiv.org/abs/2407.16607
11.12.2024 22:08 β
π 13
π 4
π¬ 0
π 0
See our latest work on (among other things) machine text detection through linguistic creativity measurement!
22.11.2024 19:07 β
π 3
π 1
π¬ 0
π 0
1/ Introducing α΄α΄α΄Ι΄κ±α΄Κα΄Κα΄Κ: a retrieval-augmented LM to help scientists synthesize knowledge π
@uwnlp.bsky.social & Ai2
With open models & 45M-paper datastores, it outperforms proprietary systems & match human experts.
Try out our demo!
openscholar.allen.ai
19.11.2024 16:30 β
π 161
π 39
π¬ 6
π 8
We'd love to be in your starter pack!
19.11.2024 01:19 β
π 2
π 0
π¬ 1
π 0