Hilary Mason: I have feelings about using the word "trained" here, as it's a LoRA'ed Qwen base model with a classification head. I would have used a different, slower approach if making it from scratch.
Hilary Mason: Overnight I trained a Jev-like model to output probability distributions over a set of 40 emoji, which runs in RAM at ~150ms per query. This was super easy and cost ~$10. It works pretty well on a …
Open the thread →