thoughts after doing a bunch of synthetic data gen for eval + environment building
thoughts after doing a bunch of synthetic data gen for eval + environment building
- LLMs are incredible projections of the world bundled into a set of weights
- but doing targeted extraction of certain distributions from those weight is
How it ranks: Backlist reads my Twitter/X timeline, scores every tweet for substance with an LLM rubric (not engagement), and publishes the daily top picks with a one-line takeaway. Curated by Surya Dantuluri.