awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
Β© 2026 Bringes Technology SRLΒ·VAT RO45896025Β·hello@awesome-repositories.com
centerforaisafety avatar

centerforaisafety/HarmBench

0
View on GitHub↗
991 stars·145 forks·Jupyter Notebook·MIT·9 viewsharmbench.org↗

HarmBench

πŸ“° Latest News πŸ“° - πŸ—‘οΈ What is HarmBench πŸ›‘οΈ - 🌐 Overview 🌐 - β˜• Quick Start β˜• - βš™οΈ Installation - πŸ› οΈ Running the Evaluation Pipeline - βž• Using your own models in HarmBench - βž• Using your own red teaming methods in HarmBench - πŸ€— Classifiers - βš“ Documentation βš“ - 🌱 HarmBench's Roadmap 🌱 -…

Features

  • Evaluation Benchmarks - Standardized framework for automated red teaming and refusal.
  • Guardrails and AI Safety - Listed in the β€œGuardrails and AI Safety” section of the The Incredible Pytorch awesome list.

Star history

Star history chart for centerforaisafety/harmbenchStar history chart for centerforaisafety/harmbench

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English β€” the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with HarmBench

These projects share indexed features with HarmBench. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • pyspur-dev/pyspurPySpur-Dev avatar

    PySpur-Dev/pyspur

    5,677View on GitHub↗
    TypeScriptagentagentsai
    View on GitHub↗5,677
  • datawhalechina/prompt-engineering-for-developersdatawhalechina avatar

    datawhalechina/prompt-engineering-for-developers

    24,267View on GitHub↗

    This project is a technical curriculum and development guide focused on large language model prompt engineering, fine-tuning, and the creation of retrieval augmented generation applications. It serves as a comprehensive resource for developers to master crafting precise instructions and textual patterns to improve the quality and predictability of model outputs. The material covers the end-to-end workflow of adapting open-source models to specific datasets and integrating language models with vector databases to generate responses based on private information. It also provides a systematic ap

    Jupyter Notebook
    View on GitHub↗24,267
  • ai45lab/openrtAI45Lab avatar

    AI45Lab/OpenRT

    257View on GitHub↗

    Open-source red teaming framework for MLLMs with 42+ attack methods

    Python
    View on GitHub↗257
  • aifeg/benchlmmAIFEG avatar

    AIFEG/BenchLMM

    86View on GitHub↗

    ECCV 2024 BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models

    Pythonbenchmarkcvdataset
    View on GitHub↗86
Compare all 30 related projects→

Frequently asked questions

What does centerforaisafety/harmbench do?

πŸ“° Latest News πŸ“° - πŸ—‘οΈ What is HarmBench πŸ›‘οΈ - 🌐 Overview 🌐 - β˜• Quick Start β˜• - βš™οΈ Installation - πŸ› οΈ Running the Evaluation Pipeline - βž• Using your own models in HarmBench - βž• Using your own red teaming methods in HarmBench - πŸ€— Classifiers - βš“ Documentation βš“ - 🌱 HarmBench's Roadmap 🌱 -…

What are the main features of centerforaisafety/harmbench?

The main features of centerforaisafety/harmbench are: Evaluation Benchmarks, Guardrails and AI Safety.

Which projects share features with centerforaisafety/harmbench?

Projects with overlapping indexed features include: pyspur-dev/pyspur. datawhalechina/prompt-engineering-for-developers β€” This project is a technical curriculum and development guide focused on large language model prompt engineering,… ai45lab/openrt β€” Open-source red teaming framework for MLLMs with 42+ attack methods. ailab-cvc/seed-bench β€” (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions. albertwy/gpt-4v-evaluation β€” Data for evaluating GPT-4V. aifeg/benchlmm β€” [ECCV 2024] BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models.