awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
facebookresearch avatar

facebookresearch/PurpleLlama

0
View on GitHub↗
4,239 Stars·740 Forks·Python·9 Aufrufe

PurpleLlama

PurpleLlama is a collection of security toolsets and frameworks designed to audit large language model vulnerabilities and implement runtime input-output guardrails. It provides a security evaluation framework and benchmark suite to quantify risks associated with prompt injections and the generation of malicious code.

The project includes a content moderator and input-output filters that use a standardized taxonomy to identify and block harmful content, jailbreaking attempts, and insecure commands. It also features capabilities for sensitive document classification to prevent the unauthorized exposure of private information.

The system covers a broad security surface including multi-language safety classification, execution-time code scanning for vulnerabilities, and cybersecurity risk benchmarking based on industry-standard frameworks.

Features

  • Content Moderation - Ships a safety model that classifies inputs and outputs to block harmful content and jailbreaking attempts.
  • Security Evaluation Suites - Provides a collection of standardized metrics and tests to assess the susceptibility of language models to security vulnerabilities.
  • Input Filters - Implements filters that analyze incoming prompts for jailbreak attempts and sensitive data before they reach the model.
  • LLM Evaluation Frameworks - Offers a set of tools and benchmarks for quantifying the risk of prompt injections and malicious code generation.
  • AI Risk Assessments - Quantifies the propensity to generate malicious code or succumb to prompt injections using industry-standard security frameworks.
  • AI Content Filters - Detects and blocks unsafe or prohibited inputs and outputs to ensure responses adhere to safety guidelines.
  • Insecure API Detection - Scans generated code for dangerous functions and insecure programming patterns to prevent malicious command execution.
  • AI Output Command Filtering - Performs security scans of generated code during execution to prevent insecure suggestions and malicious command execution.
  • LLM Prompt Injection Prevention - Identifies and blocks prompt injection and jailbreaking attempts to maintain the security and integrity of the application.
  • LLM Security Benchmarks - Provides a benchmark suite to quantify risks associated with prompt injections and malicious code generation.
  • LLM Vulnerability Assessments - Provides a comprehensive framework and benchmark suite to quantify the susceptibility of language models to security vulnerabilities.
  • Input and Output Guardrails - Provides a guardrail system that scans prompts and generated code to prevent insecure commands and data exposure.
  • LLM Security - Tests large language models for vulnerabilities to prompt injections and the generation of malicious code.
  • AI Generated Code Scanning - Inspects generated code fragments for security vulnerabilities and malicious commands prior to execution.
  • Safety Classifications - Identifies violating content across various languages using a unified taxonomy for consistent policy enforcement.
  • Data Sensitivity Classifications - Groups documents based on data sensitivity levels to prevent the unauthorized exposure of private information.
  • AI Security Frameworks - Open ecosystem for building safe and secure AI models.

Star-Verlauf

Star-Verlauf für facebookresearch/purplellamaStar-Verlauf für facebookresearch/purplellama

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu PurpleLlama

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit PurpleLlama.
  • meta-llama/purplellamaAvatar von meta-llama

    meta-llama/PurpleLlama

    4,227Auf GitHub ansehen↗

    PurpleLlama is a collection of security components and toolkits designed for large language models. It provides specialized systems including a code security scanner, a content moderation system, a prompt injection firewall, and a security assessment toolkit. The project enables the identification and blocking of jailbreaking attempts and malicious prompts during model inference. It includes capabilities for detecting violating content across multiple languages and modalities and scanning generated code for vulnerabilities to prevent the execution of insecure commands. The framework further

    Python
    Auf GitHub ansehen↗4,227
  • protectai/llm-guardAvatar von protectai

    protectai/llm-guard

    2,561Auf GitHub ansehen↗

    LLM Guard is a security firewall and guardrail framework designed to scan and sanitize inputs and outputs for large language models. It functions as a proxy gateway and security layer to block prompt injections, toxicity, and sensitive data leakage while ensuring that model interactions remain compliant with organizational policies. The system distinguishes itself through a modular scanner pipeline that utilizes local model orchestration to eliminate external network dependencies. It supports real-time security filtering via streaming chunk analysis and implements a fail-fast execution model

    Pythonadversarial-machine-learningchatgptlarge-language-models
    Auf GitHub ansehen↗2,561
  • helicone/heliconeAvatar von Helicone

    Helicone/helicone

    5,830Auf GitHub ansehen↗

    Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with large language models. By acting as a reverse-proxy, it provides a centralized layer for routing requests across multiple AI providers, allowing developers to maintain consistent application logic while gaining deep visibility into model performance, usage, and costs. The platform distinguishes itself through a robust suite of traffic management and prompt engineering tools. It enables policy-driven control, including automatic failover between providers, rate limiting, and edge-b

    TypeScript
    Auf GitHub ansehen↗5,830
  • nvidia-nemo/guardrailsAvatar von NVIDIA-NeMo

    NVIDIA-NeMo/Guardrails

    5,680Auf GitHub ansehen↗
    Pythonagentsgenerative-aiguardrails
    Auf GitHub ansehen↗5,680
Alle 30 Alternativen zu PurpleLlama anzeigen→

Häufig gestellte Fragen

Was macht facebookresearch/purplellama?

PurpleLlama is a collection of security toolsets and frameworks designed to audit large language model vulnerabilities and implement runtime input-output guardrails. It provides a security evaluation framework and benchmark suite to quantify risks associated with prompt injections and the generation of malicious code.

Was sind die Hauptfunktionen von facebookresearch/purplellama?

Die Hauptfunktionen von facebookresearch/purplellama sind: Content Moderation, Security Evaluation Suites, Input Filters, LLM Evaluation Frameworks, AI Risk Assessments, AI Content Filters, Insecure API Detection, AI Output Command Filtering.

Welche Open-Source-Alternativen gibt es zu facebookresearch/purplellama?

Open-Source-Alternativen zu facebookresearch/purplellama sind unter anderem: meta-llama/purplellama — PurpleLlama is a collection of security components and toolkits designed for large language models. It provides… protectai/llm-guard — LLM Guard is a security firewall and guardrail framework designed to scan and sanitize inputs and outputs for large… helicone/helicone — Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with… nvidia-nemo/guardrails. leondz/garak — Garak is a suite of tools for measuring AI reliability, scanning for vulnerabilities, and automating security… ibm/mcp-context-forge — mcp-context-forge is a Model Context Protocol federation gateway that unifies diverse AI tool servers and APIs into a…