BIG-Bench-Hard. Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them
568meta-prompting. Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
421dynamic-cheatsheet. Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
275hupd. The Harvard USPTO Patent Dataset
87prompt-and-rerank. Prompt-and-Rerank: A Method for Zero-Shot and Few-Shot Textual Style Transfer
36belief-in-the-machine. Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
35crowd-sampling. Follow the Wisdom of the Crowd: Effective Text Generation via Minimum Bayes Risk Decoding
20marnns. MARNNs Can Learn Generalized Dyck Languages
12lstm-eval. On Evaluating the Generalization of LSTMs in Formal Languages
9ai-as-news-intermediaries. Jupyter Notebook
3TankTouble.2.01. Java Project
1