Kristina Gligoric

Why they matter

Researcher with public evidence across AI research, Compute & infrastructure, NLP & language.

AI signals
1
past 30d
Sources
1
distinct domains
Discussions
1
past 30d
Latest signal
22d ago
View every signal from Kristina Gligoric →
Assistant Professor of Computer Science @JohnsHopkins, CS Postdoc @Stanford, PHD @EPFL, Computational Social Science, NLP, AI & Society https://kristinagligoric.com/

Articles & links

P-hacking is a long-standing problem in science, and LLMs make it worse: as tools for annotation or LLM-as-a-judge, they allow tuning parameters until a desired result appears. We propose a way to mitigate this problem, and conduct a preregistered study of its effectiveness! a…

Mitigating LLM-based p-Hacking by Preregistering for the Next LLM arxiv.org
AI Weekly's analysis
  • A new arXiv paper proposes preregistering LLM experiments and running the confirmatory analysis on the first eligible model released after registration.
  • Across 20 models from four providers and 11 configurations, the protocol blocked p-hack transfer in 73.9% and 72.7% of cases across two tasks.
  • The authors preregistered their own experiment; of 7 configurations that hacked the prior model, 6 failed to carry over to the next.
Read full analysis →
View on Bluesky · ♥ 38 ↻ 10 ↩ 2 · 3 from the directory shared this · 22d ago

Recent commentary

Attending #IC2S2!! Will be co-organizing a tutorial on Simulating Human Survey Responses with Large Language Models, today at 1.15PM, and presenting on Friday in the Understanding Large Language Models session, 10.45AM. Looking forward to catching up with old friends and meeting new ones!!

View on Bluesky · ♥ 16 ↻ 4 ↩ 1 · 17h ago

In Kristina Gligoric's orbit

Center = Kristina Gligoric. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.