Private working dataset. Conference/journal paper corpus across OpenReview venues (ICLR, NeurIPS incl. D&B/position tracks, ICML incl. position, COLM, TMLR, AISTATS, UAI, ALT, MathAI, and later additions) plus the ACL Anthology family. Coverage, per-venue availability, decisions semantics, and known biases are documented authoritatively in the GitHub repo's data/README.md — read that first; per-venue counts change as the corpus grows, so they are deliberately not duplicated here. - raw/.jsonl — canonical OpenReview snapshots, one line per - papers/ /papers.jsonl — distilled metadata (schema: data/README.md) - reviews/ /reviews.jsonl — one line per forum reply, full text - extractedtext/…
Organization
Latent Knowledge Experiments
latkes
NLP, LLMs, ML, Reasoning, Semantics, CogSci, Interpretability
Models in Library0
Datasets in Library1
Models on Hugging Face16
Followers3