HowieHwong/MemoHarness
MemoHarness: Agent Harnesses That Learn from Experience
Ph.D. student at Notre Dame | Building Frontier Learning Machine
MemoHarness: Agent Harnesses That Learn from Experience
RiskLab (ACL 2026 Demo): A Toolkit for Probing Emergent Risks in LLM-Based Multi-Agent Systems
An alignment auditing agent capable of quickly exploring alignment hypothesis
SDE-Harness (Scientific Discovery Evaluation Framework)
ProbeLLM: Automating Principled Diagnosis of LLM Failures
Arxiv Daily Agent
[ICML 2024] TrustLLM: Trustworthiness in Large Language Models
A collection of optimization problems in mathematics
A game to find putatively optimal packings in complex projective space with the goal of proving optimality of as many packings as possible.
[ICLR'24] MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use
Machine-checked Lean 4 proofs of statements from google-deepmind/formal-conjectures
A collection of formalized statements of conjectures in Lean.
[ICLR'26] Building a Foundational Guardrail for General Agentic Systems via Synthetic Data
Can We Trust Large Language Models?: A Benchmark for Responsible Large Language Models via Toxicity, Bias, and Value-alignment Evaluation
[ICLR'25] DataGen: Unified Synthetic Dataset Generation via Large Language Models
[NeurIPS'25] ChemOrch: Empowering LLMs with Chemical Intelligence via Synthetic Instructions
ValueLence: A Dashboard for In-Depth Value Probing of LLMs
Toolkit for evaluating the trustworthiness of generative foundation models.
ObscurePrompt: Jailbreaking Large Language Models via Obscure Input
Easy-to-use fine-tuning framework using PEFT (PT+SFT+RLHF with QLoRA) (LLaMA-2, BLOOM, Falcon, Baichuan)
Paper list for the survey "Combating Misinformation in the Age of LLMs: Opportunities and Challenges" and the initiative "LLMs Meet Misinformation"
Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, learderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向大型语言模型评测(例如ChatGPT、LLaMA、GLM、Baichuan等).
Related paper on the trustworthiness of large foundation models.
Paper List for ChatGPT