PhD student at UCSD studying LM reasoning and meta learning
WikiWhy. WikiWhy is a new benchmark for evaluating LLMs' ability to explain between cause-effect relationships. It is a QA dataset containing 9000+ "why" question-answer-rationale triplets.
49arc_memo. Python
46gfn_ntp. Jupyter Notebook
9cse251b-nanogpt-contest-public. Python
3llm_wrapper. Python
1geophysics_agent. Python
1mem2. Python
1