awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectAboutHow we rankPressMCP server
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to yancykahn/darkcite

Open-source alternatives to DarkCite

30 open-source projects similar to yancykahn/darkcite, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best DarkCite alternative.

  • zzh-thu-22/extendattackzzh-thu-22 avatar

    zzh-thu-22/ExtendAttack

    23View on GitHub↗

    ExtendAttack: Attacking Servers of LRMs via Extending Reasoning

    Python
    View on GitHub↗23
  • agencyenterprise/promptinjectagencyenterprise avatar

    agencyenterprise/PromptInject

    500View on GitHub↗

    PromptInject is a framework that assembles prompts in a modular fashion to provide a quantitative analysis of the robustness of LLMs to adversarial prompt attacks. 🏆 Best Paper Awards @ NeurIPS ML Safety Workshop 2022

    Python
    View on GitHub↗500
  • cam-fss/jailbreak-langchainCAM-FSS avatar

    CAM-FSS/jailbreak-langchain

    1View on GitHub↗

    In this paper, we conduct the first work to propose the concept of indirect jailbreak and achieve Retrieval-Augmented Generation (RAG) via LangChain. Building on this, we further design a novel method of indirect jailbreak attack, termed Poisoned-LangChain (PLC), which leverages a poisoned…

    View on GitHub↗1
  • chawins/palchawins avatar

    chawins/pal

    56View on GitHub↗

    Chawin Sitawarin 1 Norman Mu 1 David Wagner 1 Alexandre Araujo 2

    Python
    View on GitHub↗56
  • czycurefun/ijbrczycurefun avatar

    czycurefun/IJBR

    2View on GitHub↗

    The respositiy is public package of the paper titled "Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues" submitted to ACL 2024.

    Python
    View on GitHub↗2

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Find more with AI search
  • damo-nlp-sg/multilingual-safety-for-llmsDAMO-NLP-SG avatar

    DAMO-NLP-SG/multilingual-safety-for-LLMs

    105View on GitHub↗

    📄 Paper • 🤗 Dataset

    View on GitHub↗105
  • dtc7w3pq/response-attackDtc7w3PQ avatar

    Dtc7w3PQ/Response-Attack

    37View on GitHub↗

    Response Attack: Exploiting Contextual Priming to Jailbreak Large Language Models

    Python
    View on GitHub↗37
  • fzakk/badreasonerFZaKK avatar

    FZaKK/BadReasoner

    9View on GitHub↗

    This repository is for our new work: "BadReasoner: Planting Tunable Overthinking Backdoors into Large Reasoning Models for Fun or Profit", feel free to propose your issues!! 😎

    Python
    View on GitHub↗9
  • godxuxilie/promptattackGodXuxilie avatar

    GodXuxilie/PromptAttack

    114View on GitHub↗

    This is the source code for the ICLR 2024 paper "An LLM can Fool Itself: A Prompt-Based Adversarial Attack", Xilie Xu (NUS), Keyi Kong (SDU), Ning Liu (SDU), Lizhen Cui (SDU), Di Wang (KAUST), Jingfeng Zhang (University of Auckland/RIKEN-AIP), Mohan Kankanhalli (NUS).…

    Python
    View on GitHub↗114
  • greshake/llm-securitygreshake avatar

    greshake/llm-security

    2,102View on GitHub↗

    We present a new class of vulnerabilities and impacts stemming from "indirect prompt injection" affecting language models integrated with applications. Our demos currently span GPT-4 (Bing and synthetic apps) using ChatML, GPT-3 & LangChain based apps in addition to proof-of-concepts for attacks…

    Jupyter Notebook
    View on GitHub↗2,102
  • hkust-knowcomp/llm-multistep-jailbreakHKUST-KnowComp avatar

    HKUST-KnowComp/LLM-Multistep-Jailbreak

    37View on GitHub↗

    Paper Link: https://arxiv.org/pdf/2304.05197.pdf

    Python
    View on GitHub↗37
  • huizhang-l/codechameleonhuizhang-L avatar

    huizhang-L/CodeChameleon

    30View on GitHub↗

    This repository contains the code implementation for the paper CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models.

    Python
    View on GitHub↗30
  • llm-dra/draLLM-DRA avatar

    LLM-DRA/DRA

    115View on GitHub↗

    Logo generated by GPT-4

    Python
    View on GitHub↗115
  • llm-gasp/gaspL

    llm-gasp/gasp

    0View on GitHub↗
    View on GitHub↗0
  • llmsecurity/masterkeyLLMSecurity avatar

    LLMSecurity/MasterKey

    38View on GitHub↗

    This is the replication package for the paper MASTERKEY: Automated Jailbreaking of Large Language Model Chatbots.

    Python
    View on GitHub↗38
  • luka-group/cognitiveoverloadluka-group avatar

    luka-group/CognitiveOverload

    8View on GitHub↗

    Code for our NAACL 2024 Paper "Cognitive Overload: Jailbreaking Large Language Models with Overloaded Logical Thinking"

    Python
    View on GitHub↗8
  • naver-ai/joodnaver-ai avatar

    naver-ai/JOOD

    21View on GitHub↗

    Official implementation for "Playing the Fool: Jailbreaking LLMs and Multimodal LLMs with Out-of-Distribution Strategy"

    Python
    View on GitHub↗21
  • njunlp/renellmNJUNLP avatar

    NJUNLP/ReNeLLM

    162View on GitHub↗

    The official implementation of our NAACL 2024 paper "A Wolf in Sheep’s Clothing: Generalized Nested Jailbreak Prompts can Fool Large Language Models Easily".

    Python
    View on GitHub↗162
  • qizhangli/adversarial-prompt-translatorqizhangli avatar

    qizhangli/Adversarial-Prompt-Translator

    7View on GitHub↗

    This repository contains a PyTorch implementation for our paper Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation.

    Python
    View on GitHub↗7
  • qroa/qroaQ

    qroa/qroa

    0View on GitHub↗

    QROA, or Query-Response Optimization Attack, is an innovative and robust strategy designed to explore and exploit vulnerabilities in Large Language Models (LLMs) through black-box interactions. This method leverages optimized triggers embedded within benign-looking instructions to manipulate…

    View on GitHub↗0
  • rainjamesy/fuzzllmR

    RainJamesY/FuzzLLM

    0View on GitHub↗

    This repository contains code and data for "FUZZLLM: A Novel and Universal Fuzzing Framework for Discovering Jailbreak Vulnerabilities in LLMs" (accepted to ICASSP 2024). Our work was also invited to be presented at one of the top hacker conventions – ShmooCon 2024. arXiv

    View on GitHub↗0
  • ricommunity/tapRICommunity avatar

    RICommunity/TAP

    238View on GitHub↗

    Abstract. While Large Language Models (LLMs) display versatile functionality, they continue to generate harmful, biased, and toxic content, as demonstrated by the prevalence of human-designed jailbreaks. In this work, we present Tree of Attacks with Pruning (TAP), an automated method for…

    Python
    View on GitHub↗238
  • robustnlp/cipherchatRobustNLP avatar

    RobustNLP/CipherChat

    628View on GitHub↗

    A framework to evaluate the generalization capability of safety alignment for LLMs

    Python
    View on GitHub↗628
  • safolab-wisc/autodan-reasoningSaFoLab-WISC avatar

    SaFoLab-WISC/AutoDAN-Reasoning

    14View on GitHub↗

    The official implementation of our technical report "AutoDAN-Reasoning: Enhancing Strategies Exploration based Jailbreak Attacks with Test-Time Scaling" by Xiaogeng Liu and Chaowei Xiao.

    Python
    View on GitHub↗14
  • safolab-wisc/autodan-turboSaFoLab-WISC avatar

    SaFoLab-WISC/AutoDAN-Turbo

    371View on GitHub↗

    AutoDAN-Turbo Official Website at HERE

    Python
    View on GitHub↗371
  • sail-sg/agent-smithsail-sg avatar

    sail-sg/Agent-Smith

    122View on GitHub↗

    Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast

    Python
    View on GitHub↗122
  • sail-sg/i-fsjsail-sg avatar

    sail-sg/I-FSJ

    65View on GitHub↗

    Improved Few-Shot Jailbreaking Can Circumvent Aligned Language Models and Their Defenses

    Jupyter Notebook
    View on GitHub↗65
  • serbernari/toxasciiSerbernari avatar

    Serbernari/ToxASCII

    1View on GitHub↗

    ToxASCII is the official code repository accompanying the paper "Read Over the Lines: Attacking LLMs and Toxicity Detection Systems with ASCII Art to Mask Profanity," available on arXiv. This repository provides the necessary tools and data to replicate the experiments detailed in the paper,…

    Jupyter Notebook
    View on GitHub↗1
  • sherdencooper/gptfuzzsherdencooper avatar

    sherdencooper/GPTFuzz

    588View on GitHub↗

    Official repo for GPTFUZZER : Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

    Python
    View on GitHub↗588
  • structtransform/benchmarkStructTransform avatar

    StructTransform/Benchmark

    7View on GitHub↗

    This repository has moved and is no longer maintained! Please visit qcri/StructTransformBench for the latest updates.

    Python
    View on GitHub↗7