Maria Antoniak

Assistant professor of CS, NLP and cultural analytics

Why they matter

Assistant professor of CS, NLP and cultural analytics with public evidence across AI research, Culture, work & education, NLP & language.

AI signals
21
past 30d
Sources
17
distinct domains
Discussions
35
past 30d
Latest signal
1d ago
View every signal from Maria Antoniak →
NLP, cultural analytics. Assistant Professor of CS at University of Colorado Boulder. Books, bikes, games, art. https://maria-antoniak.github.io

Articles & links

Maria Antoniak reposted
@davidjbianco.bsky.social

HuggingFace got hacked by an AI. What stuck out to me was the guardrail asymmetry. The attacker had no constraints, but HF's response ran afoul of the abuse guardrails, forcing them into an unplanned switch to local models. Another aspect for your IR plans. huggingface.co/blog…

Security incident disclosure — July 2026 huggingface.co
AI Weekly's analysis
  • The attacker's agent ran 17,000+ actions across short-lived sandboxes, compressing multi-stage lateral movement into a single weekend.
  • Commercial model APIs blocked forensic requests containing real exploit artifacts, forcing Hugging Face to pivot to open-weight GLM 5.2 on private infrastructure.
  • The intrusion entered via a remote dataset RCE loader and configuration template injection, not through the model-serving layer.
Read full analysis →
View on Bluesky →
Maria Antoniak reposted
Juan Diego Rodriguez @juand-r.bsky.social

2) Characterizing Narrative Content in Web-scale LLM Pretraining Data, by @teagrjohnson.bsky.social, @elliottash.bsky.social, @andrewpiper.bsky.social, @mariaa.bsky.social arxiv.org/abs/2606.19468 Why: annotation and analysis of narrative features across the pretraining data (…

Characterizing Narrative Content in Web-scale LLM Pretraining Data arxiv.org
AI Weekly's analysis
  • A new arXiv preprint introduces NarraBERT, a RoBERTa-based classifier, and applies it to 3 million passages from the 3-trillion-token Dolma corpus.
  • The framework operationalizes three narrative elements, agency, setting, and events, across 11 interpretable dimensions, trained on 400 annotated passages.
  • The authors report narrative qualities are unequally distributed across pretraining sources and topics in ways current curation practices do not measure.
Read full analysis →
View on Bluesky →
Maria Antoniak reposted
Anjalie Field @anjalief.bsky.social

New press release about our recently released COLM paper, work by Katherine Van Koevering! hub.jhu.edu/2026/09/21/a... Check out the full paper here: arxiv.org/abs/2608.13328

It's How You Ask: Gender-Associated Linguistic Bias in LLMs arxiv.org
AI Weekly's analysis
  • Prompts using hedges, tag questions, and collective reference drew shorter, less sophisticated, less formal LLM replies across three document types and four models.
  • Linguistic register outweighed explicit cues like sign-off names in shaping response quality, according to the paper's representational analysis.
  • The authors report the effect is encoded in early transformer layers and entangled with other features, making post-hoc mitigation difficult.
Read full analysis →
View on Bluesky →
Maria Antoniak reposted
Lauren Klein @laurenfklein.bsky.social

After far too many AI-generated final projects, and my own uncertainty about what students really need to know about NLP these days, I decided to overhaul my Text as Data class so that it's structured around a single project: training up a C19 LLM from scratch. Here we go! git…

GitHub - laurenfklein/dsci340-fall26: Notebooks and other course materials for Emory DSCI 340 (Fall 2026 - Klein) github.com
AI Weekly's analysis
Read full analysis →
View on Bluesky →
Maria Antoniak reposted
Mark Riedl @markriedl.bsky.social

Nobody know what consumer device OpenAI is making, but Apple says OpenAI stole trade secrets to do whatever it is www.cnbc.com/amp/2026/07/...

Apple sues OpenAI alleging trade secret theft cnbc.com
AI Weekly's analysis
  • Apple filed Friday in the Northern District of California, accusing OpenAI of trade-secret misappropriation and breach of contract tied to hardware development.
  • The complaint names OpenAI hardware chief Tang Tan, a former Apple VP of product design, and ex-Apple engineer Chang Liu, who allegedly kept his Apple laptop.
  • OpenAI's $6.4 billion acquisition of Jony Ive's startup IO Products, also named as a defendant, sits at the center of the dispute.
Read full analysis →
View on Bluesky →

Our paper: arxiv.org/abs/2606.22748 Wired article:

AI Fiction in the Wild arxiv.org
AI Weekly's analysis
  • A new arXiv preprint analyzing over 500,000 anonymized English-language ChatGPT conversations finds more than one third involve fiction generation.
  • The behavior is dominated by power users the authors call 'infinite story demanders,' who repeatedly request and revise variations of similar narratives.
  • Users gravitate especially toward fanfiction and erotica, favoring generic forms, repetition, immediacy, and niche combinations of story elements.
Read full analysis →
View on Bluesky · ♥ 7 ↻ 1 ↩ 0 · 2 from the directory shared this · 7d ago

Recent commentary

Is AI useful... Well I've avoided Figma for years because every time I opened it, it looked too annoying to learn. But now they have a chatbot, and I made three websites yesterday.

View on Bluesky · ♥ 80 ↻ 1 ↩ 1 · 49d ago

People working in AI industry posting nonsense on X are not “putting their careers on the line.” They’ve got a ton of $$ and will be fine. A lot of this discourse is just weird in-group social signaling explicitly *for* career advancement. Genuine concerns exist but good luck disentangling all that.

View on Bluesky · ♥ 63 ↻ 9 ↩ 0 · 13d ago

I’m on my way to my first #IC2S2! 🌧️ I’ll be supporting presentations about computational narrative analysis by my students and mentees @teagrjohnson.bsky.social and @umagunturi1.bsky.social. If you’re interested in storytelling, pretraining data, social media, or language models, come talk to us!

View on Bluesky · ♥ 32 ↻ 3 ↩ 1 · 57d ago

I'm looking for examples of successful NSF CAREER proposals and budgets, especially those written by NLP researchers. If you would be willing to share your successful proposal, or if you know of successful proposals that others have shared publicly, please let me know!

View on Bluesky · ♥ 12 ↻ 2 ↩ 1 · 70d ago

I'm looking for 1-2 emergency reviewers for a #COLM2026 paper on LLMs and science. Paper looks like a nice read to me! Reply or message me if you might be able to write a review by end of week.

View on Bluesky · ♥ 6 ↻ 2 ↩ 1 · 132d ago

In Maria Antoniak's orbit

Center = Maria Antoniak. Left = members they follow (green edges). Right = members who follow them (blue edges). Top = mutual follows (orange edges, slightly larger). Drag any node to reposition; click to open that profile.

Are you Maria Antoniak? Show it.

Add the Who’s Who of AI badge to your site or bio. It links back to this profile.

Listed in AI Weekly's Who's Who of AI

Markdown: [![Listed in AI Weekly's Who's Who of AI](https://aiweekly.co/modules/custom/aiweekly_whoswho/images/whoswho-badge.svg)](https://aiweekly.co/whos-who/person/maria-antoniak)