Skip to content
#

llm-safety

Here are 402 public repositories matching this topic...

[ICML 2024] TrustLLM: LLM trustworthiness evaluation across truthfulness, safety, fairness, robustness, privacy and ethics. Python/CLI toolkit for local models, compatible APIs and AI agent orchestration.

  • Updated Oct 4, 2026
  • Python

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses | 500+ Papers | Perception, Cognition, Planning, Interaction, Agentic System

  • Updated Oct 5, 2026
  • Python
awesome-llm-agent-papers

A curated, continuously updated reading list of 500+ papers on LLM agents: planning, memory, tool use, multi-agent, evaluation & safety. Companion to the survey 'LLM Agents: A Survey'.

  • Updated Oct 3, 2026
  • Python

Add this topic to your repo

To associate your repository with the llm-safety topic, visit your repo's landing page and select "manage topics."

Learn more