awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
allenai avatar

allenai/wildteaming

0
View on GitHub↗
43 stars·5 forks·Python·7 views

Wildteaming

Authors: Liwei Jiang, Kavel Rao ⭐, Seungju Han ⭐, Allyson Ettinger, Faeze Brahman, Sachin Kumar, Niloofar Mireshghallah, Ximing Lu, Maarten Sap, Yejin Choi, Nouha Dziri       ⭐ Co-second authors

Features

  • Evaluation Benchmarks - Framework for scaling in-the-wild red teaming and safety.

Star history

Star history chart for allenai/wildteamingStar history chart for allenai/wildteaming

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with Wildteaming

These projects share indexed features with Wildteaming. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • pyspur-dev/pyspurPySpur-Dev avatar

    PySpur-Dev/pyspur

    5,677View on GitHub↗
    TypeScriptagentagentsai
    View on GitHub↗5,677
  • datawhalechina/prompt-engineering-for-developersdatawhalechina avatar

    datawhalechina/prompt-engineering-for-developers

    24,267View on GitHub↗

    This project is a technical curriculum and development guide focused on large language model prompt engineering, fine-tuning, and the creation of retrieval augmented generation applications. It serves as a comprehensive resource for developers to master crafting precise instructions and textual patterns to improve the quality and predictability of model outputs. The material covers the end-to-end workflow of adapting open-source models to specific datasets and integrating language models with vector databases to generate responses based on private information. It also provides a systematic ap

    Jupyter Notebook
    View on GitHub↗24,267
  • ai45lab/openrtAI45Lab avatar

    AI45Lab/OpenRT

    257View on GitHub↗

    Open-source red teaming framework for MLLMs with 42+ attack methods

    Python
    View on GitHub↗257
  • aifeg/benchlmmAIFEG avatar

    AIFEG/BenchLMM

    86View on GitHub↗

    ECCV 2024 BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models

    Pythonbenchmarkcvdataset
    View on GitHub↗86
Compare all 30 related projects→

Frequently asked questions

What does allenai/wildteaming do?

Authors: Liwei Jiang, Kavel Rao ⭐, Seungju Han ⭐, Allyson Ettinger, Faeze Brahman, Sachin Kumar, Niloofar Mireshghallah, Ximing Lu, Maarten Sap, Yejin Choi, Nouha Dziri       ⭐ Co-second authors

What are the main features of allenai/wildteaming?

The main features of allenai/wildteaming are: Evaluation Benchmarks.

Which projects share features with allenai/wildteaming?

Projects with overlapping indexed features include: pyspur-dev/pyspur. datawhalechina/prompt-engineering-for-developers — This project is a technical curriculum and development guide focused on large language model prompt engineering,… ai45lab/openrt — Open-source red teaming framework for MLLMs with 42+ attack methods. ailab-cvc/seed-bench — (CVPR2024)A benchmark for evaluating Multimodal LLMs using multiple-choice questions. albertwy/gpt-4v-evaluation — Data for evaluating GPT-4V. aifeg/benchlmm — [ECCV 2024] BenchLMM: Benchmarking Cross-style Visual Capability of Large Multimodal Models.