I think there is maybe some interplay between LLM architecture, RL and instruct data that is missing. Right now L…
I think there is maybe some interplay between LLM architecture, RL and instruct data that is missing. Right now LLM "processing depth" seems really shallow, reasoning traces look like retrieval + incremental reasoning, which is not what one
How it ranks: Backlist reads my Twitter/X timeline, scores every tweet for substance with an LLM rubric (not engagement), and publishes the daily top picks with a one-line takeaway. Curated by Surya Dantuluri.