Skip to content
@AI-secure

AI Secure

UIUC Secure Learning Lab

Popular repositories Loading

  1. DecodingTrust DecodingTrust Public

    A Comprehensive Assessment of Trustworthiness in GPT Models

    Python 317 62

  2. AgentPoison AgentPoison Public

    [NeurIPS 2024] Official implementation for "AgentPoison: Red-teaming LLM Agents via Memory or Knowledge Base Backdoor Poisoning"

    Python 244 35

  3. DBA DBA Public

    DBA: Distributed Backdoor Attacks against Federated Learning (ICLR 2020)

    Python 206 49

  4. Certified-Robustness-SoK-Oldver Certified-Robustness-SoK-Oldver Public

    This repo keeps track of popular provable training and verification approaches towards robust neural networks, including leaderboards on popular datasets and paper categorization.

    97 10

  5. DecodingTrust-Agent DecodingTrust-Agent Public

    Python 92 13

  6. RedCode RedCode Public

    [NeurIPS'24] RedCode: Risky Code Execution and Generation Benchmark for Code Agents

    Python 91 12

Repositories

Showing 10 of 60 repositories
  • AgentPoison Public

    [NeurIPS 2024] Official implementation for "AgentPoison: Red-teaming LLM Agents via Memory or Knowledge Base Backdoor Poisoning"

    AI-secure/AgentPoison's past year of commit activity
    Python 244 MIT 35 2 0 Updated Oct 6, 2026
  • RedCode Public

    [NeurIPS'24] RedCode: Risky Code Execution and Generation Benchmark for Code Agents

    AI-secure/RedCode's past year of commit activity
    Python 91 MIT 12 0 0 Updated Aug 19, 2026
  • SafeAuto Public

    [ICML 2025] SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models

    AI-secure/SafeAuto's past year of commit activity
    Python 28 2 6 0 Updated Aug 2, 2026
  • AI-secure/DecodingTrust-Agent's past year of commit activity
    Python 92 Apache-2.0 13 3 0 Updated Jun 18, 2026
  • ShieldNet Public
    AI-secure/ShieldNet's past year of commit activity
    JavaScript 0 0 0 0 Updated Mar 16, 2026
  • UDora Public

    [ICML 2025] UDora: A Unified Red Teaming Framework against LLM Agents

    AI-secure/UDora's past year of commit activity
    Python 39 7 1 0 Updated Jun 24, 2025
  • PolyGuard Public
    AI-secure/PolyGuard's past year of commit activity
    Python 24 3 3 0 Updated Jun 18, 2025
  • AdvAgent Public
    AI-secure/AdvAgent's past year of commit activity
    Jupyter Notebook 26 0 6 0 Updated May 28, 2025
  • MMDT Public

    Comprehensive Assessment of Trustworthiness in Multimodal Foundation Models

    AI-secure/MMDT's past year of commit activity
    Jupyter Notebook 29 2 1 0 Updated Mar 15, 2025
  • aug-pe Public

    [ICML 2024 Spotlight] Differentially Private Synthetic Data via Foundation Model APIs 2: Text

    Python 61 Apache-2.0 20 1 0 Updated Jan 11, 2025