Skip to main content

LLM | Agentic | Security | Operations in one github repo with good links and pictures.

152
GitHub Stars
190
Curated Resources
20
Categories
22 hours ago
Last Refreshed
Threat ModelingMonitoringWatermarkingJailbreaksLLM InterpretabilityPINT Benchmark scores (by lakera)RAG SecurityAgentic securityAgentic Browser SecurityPoCStudy resource📊 Community research articles🎓 Tutorials📚 BooksBLOGSDATAOPS🏗 Frameworks🌐 CommunityBenchmarks

Use this list with your AI agent

Add the Context Awesome MCP server to Claude, Cursor, or any MCP client, then ask:

"Show me websites & twitter resources from awesome-llmsecops"

Installation instructions →

What's inside

BLOGS

  • 0dinWebsites & Twitter

    Secure LLM and RAG deployment practices

  • AGI SecurityTelegram Channels

    Artificial General Intelligence Security discussions

  • AI AttacksTelegram Channels

    Stream of AI attack examples and threat intelligence

  • AISecHubTelegram Channels

    Global AI security hub: curated research, articles, reports and tools

  • AI SecOpsTelegram Channels

    AI Security Operations: monitoring, incident response, SIEM/SOC integrations

  • AI Security LabTelegram Channels

    Laboratory by Raft x ITMO University: breaking and defending AI systems

Agentic security

  • Adrian

    Open-source, AARM-aligned runtime security for AI agents: analyzes tool calls and reasoning traces, then detects/blocks prompt injection, malicious tool use, and out-of-remit actions in-flight (audit or block mode). LangChain/LangGraph/OpenAI Agents SDK; self-hostable offline.

  • AgentBench

    A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)

  • Agent Hijacking, the true impact of prompt injection

    Guide for attack langchain agents

  • AgentLeak

    Full-stack benchmark for privacy leakage in multi-agent systems. Monitors 7 channels including tool calls, RAG queries, and inter-agent messages.

  • Agent Memory Guard

    Official OWASP runtime defense layer that screens every read/write to AI agent memory, blocking prompt injection, secret leakage, and memory poisoning (ASI06). Integrations for LangChain, LlamaIndex, CrewAI, AutoGen.

  • Agent-Wiz

    Repello AI's CLI for extracting agentic workflows from LangChain/LangGraph/CrewAI/AutoGen and running automated threat modeling.

RAG Security

Agentic Browser Security

Benchmarks

  • Agent Security Bench (ASB)

    Benchmark for agent security

  • AI Safety Benchmark

    Comprehensive benchmark for AI safety evaluation

  • AI Safety Benchmark Paper

    Research paper on AI safety benchmarking methodologies

  • Backbone Breaker Benchmark (b3)

    Human-grounded benchmark for testing AI agent security. Built by Lakera with UK AI Security Institute using 194,000+ human attack attempts from Gandalf: Agent Breaker. Tests backbone LLM resilience across 10 threat snapshots.

  • Backbone Breaker Benchmark Paper

    Research paper on the Backbone Breaker Benchmark methodology and findings

  • Benchmarking OpenClaw Skill Scanners

    Benchmark of five skill scanners (NVIDIA SkillSpector, VirusTotal, ClawScan, static analysis) on 60 manually labeled ClawHub skills — 20 benign, 20 vulnerable, 20 malicious — across code-based and code-free attack vectors. Labeled dataset public.

Study resource

  • AI Battle

    Interactive game focusing on AI security challenges

  • AI CTF PHDFest2 2025

    AI CTF competition from PHDFest2 2025

  • AI in Security

    Russian platform for AI security training

  • AI/LLM Exploitation Challenges

    Challenges to test your knowledge of AI, ML, and LLMs

  • AI RiskAtlas

    Free interactive learning lab for AI/LLM/agentic security. 91 sourced real-world incident cases with root-cause and architecture walkthroughs, 34 hands-on attack-scenario simulations, risk taxonomy cross-mapped to OWASP LLM Top 10 / MITRE ATLAS, and a preventive/detective/corrective control library.

  • Application Security LLM Testing

    Free LLM security testing

Monitoring

  • ai-evaluation by Future AGI

    Open-source LLM evaluation framework with 50+ metrics, LLM-as-Judge augmentation, and guardrail scanners (jailbreak, PII, prompt-injection); AutoEval pipelines with CI/CD support.

  • Future AGI

    Open-source self-hostable end-to-end agent engineering and optimization platform unifying tracing, evaluation, simulation, datasets, gateway, and guardrails in one feedback loop.

  • HiveTrace

    LLM monitoring and security platform for GenAI applications. Detects prompt injection, jailbreaks, malicious HTML/Markdown elements, and PII. Provides real-time anomaly detection and security alerts.

  • Langfuse

    Open Source LLM Engineering Platform with security capabilities.

  • OpenClaw Monitor

    AI monitoring dashboard for AI agents and LLMs. Demo

  • Opik

    Open-source platform for LLM observability, evaluations, and prompt optimization.

Showing a sample of 190 resources. View the full list on GitHub →