AI Safety & Alignment
Showing 47 of 47 AI Safety & Alignment tools on page 1 of 1.
AI Safety & Alignment results
A4-Artificial-Intelligence
this is a repo regarding the ethics of AI. Contribute to AlivadTheImpala/A4-Artificial-Intelligence development by…
Ai_Ethics
Data Science - Ethical Artificial Intelligence (SS 2022) - FH-kiel-lectures/AI_Ethics
Ai-Ethics
AI Ethics Committee at PARC. Contribute to PARC/ai-ethics development by creating an account on GitHub.
Ai-Ethics
Ethical Guidelines for Human-Advanced Intelligence Interaction - cbhanni/AI-Ethics
Ai-Ethics-Experiments
Analyzing AI's Responses To Ethical Dilemmas. Contribute to jtrugman/ai-ethics-experiments development by creating…
Ai-Ethics-Fairness-And-Bias
Sample project using IBM's AI Fairness 360 is an open source toolkit for determining, examining, and mitigating…
Ai-Ethics-Framework
Ethical questions, risks and issues to think about to help create responsible AI products and services -…
Ai-Ethics-Tool-Landscape
AI Ethics Tool Landscape. Contribute to EdwinWenink/ai-ethics-tool-landscape development by creating an account on…
Ai-Ethics-Toolkit-For-Kenya
The AI Ethics Toolkit for Kenya is an open-source project aimed at building a comprehensive set of tools and…
AI Fairness 360
A comprehensive set of fairness metrics for datasets and machine learning models, explanations for these metrics,…
AI Risks that Could Lead to Catastrophe | CAIS
There are many potential risks from AI. CAIS focusses on mitigating risks that could lead to catastrophic outcomes…
Ai-Safety-Ethics
Bridging research problems of the fields AI safety and AI ethics - gyevnarb/ai-safety-ethics
Ai-Unexpected-Behaviors
catalog of unexpected behaviors of ai which can turn out to be failures that compromise the initial objectives,…
Aiethicsconsulting
Provides expert consulting services on AI ethics, helping organizations navigate complex ethical issues in AI. -…
Alignment Handbook
Robust recipes to align language models with human and AI preferences - huggingface/alignment-handbook
Anthropic Transparency Hub
A look at Anthropic's key processes, programs, and practices for responsible AI development.
Biasbounty1_Humaneintelligence
This repo contains my coding notebook for the tutorial series I made for the beginner level bias bounty challenge…
Calpoly-Aiel
A site that hosts the Cal Poly AI Ethics Lab. Contribute to Cal-Poly-AIEL/calpoly-aiel development by creating an…
Center for AI Safety
Center for AI Safety. Reducing societal-scale risks from AI by advancing safety research, building the field of AI…
CogBias AI
See hidden influences to reveal the truth with our An AI Platform for the detection of cognitive biases. Our…
Corporate_Ai_Ethics_Guideline_Analysis
corporate AI ethics guidelines analysis
Derai
Our Digital Ethics and Responsible AI code reorganizes information in predictions and user behavior to prioritize…
Dilemmasearcher
Project in INFO381 - Advanced Topics in Artificial Intelligence: Moral and Ethics in AI, Spring 2017 -…
Ella
ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment - TencentQQGYLab/ELLA
Ethics-And-Ai
Articulate Rise tabs intraction on Ethics and AI. Contribute to herogrl/ethics-and-ai development by creating an…
Ethics-In-Ai
Ethical issues existing in the AI systems. Contribute to tilwani/Ethics-in-AI development by creating an account on…
Ethicsai
Ethics in the Age of AI. Contribute to dennislamcv1/EthicsAI development by creating an account on GitHub.
Ethicsai
A Study on the Ethics of AI. Contribute to prodp/EthicsAI development by creating an account on GitHub.
Ethicscore
Ethics-by-design AI applications evaluator and enabler. - natalia-moral/ethicscore
Face-Recognition
When I did this project (over 3 years ago), my intentions were to use it to nab criminals. With awareness of ethics…
Fate-In-Ai
Relevant information about the project on Fairness, Accountability, Transparency and Ethics in Artificial…
Hamid
AI Ethics for Future . Contribute to HamidurRahmanBd/Hamid development by creating an account on GitHub.
HarmBench
HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal -…
Heretic
Fully automatic censorship removal for language models - p-e-w/heretic
Llm-Warden
A simple jailbreak detection tool for safeguarding LLMs. - jackhhao/llm-warden
Mural AI
Unlock creative solutions faster with Mural AI. Empower your teams to innovate, align, and collaborate efficiently…
NeMo Guardrails
NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational…
Nottsai-Meetup-June4-2019
Ethics Guidelines for Trustworthy AI. Contribute to Lazymindz/nottsai-meetup-june4-2019 development by creating an…
OsherAI
We map, automate and maintain the manual processes slowing your business down, using automation, RPA and ethical AI.…
Pittchallenge2023
Governance and Ethics in AI - Pitt Challenge.
Reddit_Ai_Topic_Analysis
Unveiling Trends in AI Ethics: Exploring the Ethical Dimensions of AI, Prepared for SICSS 2023 - Tor Vergata -…
Relaieo
Relational AI Ethics Ontology. Contribute to Audit4SG/RelAIEO development by creating an account on GitHub.
Responsible AI Toolbox
Responsible AI Toolbox is a suite of tools providing model and data exploration and assessment user interfaces and…
responsible.ai
The Responsible AI Institute accelerates trustworthy AI through standards-aligned certification, governance…
RiskAssessmentAI
Automates the completion of security questionnaires received from customers and prospects.
The Alignment Problem
Brian Christian. AI safety explained for general audiences.
Weavesphere-2022-Ai-Ethics
AI Ethics: The Content Design Perspective. Contribute to spackows/WEAVESPHERE-2022-AI-Ethics development by creating…