Perspective is an API that uses machine learning models to score the perceived impact a comment might have on a conversation. See https://developers.perspectiveapi.com for more information.
A framework for Privacy Preserving Machine Learning
A Python package to assess and improve fairness of machine learning models.
📰 Latest News 📰 - 🗡️ What is HarmBench 🛡️ - 🌐 Overview 🌐 - ☕ Quick Start ☕ - ⚙️ Installation - 🛠️ Running the Evaluation Pipeline - ➕ Using your own models in HarmBench - ➕ Using your own red teaming methods in HarmBench - 🤗 Classifiers - ⚓ Documentation ⚓ - 🌱 HarmBench's Roadmap 🌱 -…
The main features of anthropics/constitutional-ai are: Guardrails and AI Safety.
Open-source alternatives to anthropics/constitutional-ai include: conversationai/perspectiveapi — Perspective is an API that uses machine learning models to score the perceived impact a comment might have on a… facebookresearch/crypten — A framework for Privacy Preserving Machine Learning. fairlearn/fairlearn — A Python package to assess and improve fairness of machine learning models. guardrails-ai/guardrails — Guardrails is a Python SDK that wraps calls to large language models with configurable validation pipelines,… interpretml/interpret — Interpret is an interpretable machine learning library and glassbox model framework. It provides toolkits for training… centerforaisafety/harmbench — 📰 Latest News 📰 - 🗡️ What is HarmBench 🛡️ - 🌐 Overview 🌐 - ☕ Quick Start ☕ - ⚙️ Installation - 🛠️ Running the…