Yue Huang

@HowieHwong · User

GitHub profile ↗ · Compare

Ph.D. student at Notre Dame | Building Frontier Learning Machine

University of Notre DameBay Area, USA95 followers27 repositories

Repositories

HowieHwong/RiskLab

RiskLab (ACL 2026 Demo): A Toolkit for Probing Emergent Risks in LLM-Based Multi-Agent Systems

★ 33PythonForks 1

HowieHwong/ProbeLLM

ProbeLLM: Automating Principled Diagnosis of LLM Failures

★ 18PythonForks 1

HowieHwong/TrustLLM

[ICML 2024] TrustLLM: Trustworthiness in Large Language Models

★ 633PythonForks 68

HowieHwong/GameofSloanes

A game to find putatively optimal packings in complex projective space with the goal of proving optimality of as many packings as possible.

★ 0Forks 0

HowieHwong/MetaTool

[ICLR'24] MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use

★ 120PythonForks 12

HowieHwong/Agentic-Guardian

[ICLR'26] Building a Foundational Guardrail for General Agentic Systems via Synthetic Data

★ 49PythonForks 6

HowieHwong/TrustGPT

Can We Trust Large Language Models?: A Benchmark for Responsible Large Language Models via Toxicity, Bias, and Value-alignment Evaluation

★ 25PythonForks 5

HowieHwong/DataGen

[ICLR'25] DataGen: Unified Synthetic Dataset Generation via Large Language Models

★ 69PythonForks 4

HowieHwong/ChemOrch

[NeurIPS'25] ChemOrch: Empowering LLMs with Chemical Intelligence via Synthetic Instructions

★ 14PythonForks 2

HowieHwong/llm-misinformation-survey

Paper list for the survey "Combating Misinformation in the Age of LLMs: Opportunities and Challenges" and the initiative "LLMs Meet Misinformation"

★ 1Forks 0

HowieHwong/Awesome-LLM-Eval

Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, learderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向大型语言模型评测(例如ChatGPT、LLaMA、GLM、Baichuan等).

★ 0Forks 0