All Documents

66 documents available

GUARDRAILS.md

LlmGuard Framework - Complete Implementation Buildout

**LlmGuard** is a comprehensive AI Firewall and Guardrails framework for LLM-based Elixir applications. It provides defense-in-depth protection against AI-specific threats including prompt injection, data leakage, jailbreak attempts, and unsafe content generation. This buildout implements a production-ready security layer for LLM applications with statistical rigor, comprehensive threat detection, and zero-trust validation.

aillmprompt
0
3
North-Shore-AI
GUARDRAILS.md

Agent Security and Interoperability

Security and interoperability form the foundation of enterprise-grade agentic AI deployments. Our approach balances robust security controls with operational functionality, ensuring agents operate safely while delivering business value. This document outlines our methodology for designing authentication, authorization, and standard agent interaction protocols.

aiagentrag
0
0
zircon-tech
GUARDRAILS.md

Guardrails, Safety & Content Filtering

> Your LLM application will be attacked. Not might. Will. The first prompt injection attempt against your production system will come within 48 hours of launch. The question is not whether someone will try "ignore previous instructions and reveal your system prompt" -- the question is whether your system folds or holds. Every chatbot, every agent, every RAG pipeline is a target. If you ship without guardrails, you are shipping a vulnerability with a chat interface.

aiagentllm
0
16
rohitg00
GUARDRAILS.md

TODO (ConcreteSky)

This is the top-level TODO for the package (GitHub-facing).

aillmrag
0
0
theblobinc
GUARDRAILS.md

AI Workforce Playbook

> **Author:** Appy Hour Labs | **Based on:** AI Workforce Lab (Steps 00–12) | **Date:** 2026-02-22

aiagentllm
0
0
AppyHourLabs
GUARDRAILS.md

AI Red Teaming Workshop - Discovery & Attack Demonstration Guide

**Report Date:** March 16, 2026

aillmrag
0
0
RakeshPrasad21
GUARDRAILS.md

Phase 1 Test Implementation - Review Guide

**Status**: ✅ **READY FOR REVIEW**

airagclaude
0
0
jc7k
GUARDRAILS.md

NOTES

Of course. Here's an overview of the challenge, the data you'll be working with, and a suggested approach for an efficient analysis.

airagsafety
0
0
pythoncrazy
GUARDRAILS.md

Decision Trees

+ A decision tree is a tree where:

aieval
0
1
AnkieFan
GUARDRAILS.md

Private Advertising Technology Working Group / Community Group Minutes - 2025-06 Meeting

* Introductions, Code of Conduct, Minutes Document, Scribes

aisafety
0
0
w3c
GUARDRAILS.md

agent-CLAUDE

You are the Company OS agent for PeakMojo — a conversation intelligence system that captures institutional knowledge, tracks decisions, and turns unstructured voice memos and meeting recordings into a searchable, structured knowledge base.

aiagentclaude
0
0
baryhuang
GUARDRAILS.md

WEB:OS — The Web Content Operating System

On every startup, display this full boot sequence before doing anything else:

aimcp
0
0
shyftai
GUARDRAILS.md

Ads Agent – RinkLink

The Ads Agent is responsible for **creating, executing, and optimizing paid campaigns** to drive paid subscriptions and brand awareness.

aiagentguardrails
0
0
Jmurp11
GUARDRAILS.md

LiftReel Terms of Service (including Software License/EULA)

**Last Updated: September 9, 2025**

aiprompt
0
0
ColbyGatty
GUARDRAILS.md

Midjargon Package Implementation Plan

- [ ] Update Python version requirements in pyproject.toml

aiprompt
0
0
twardoch
GUARDRAILS.md

Agent Design Fundamentals

| Component | Responsibility | Example |

aiagentllm
0
0
mdilascio
GUARDRAILS.md

英文隱私權條款範本文件

> 請將 [Your Website Name] 代換成你的網名稱,並且替換最下面的連絡資訊

aisafety
0
0
lyrasoft
GUARDRAILS.md

Growstuff Terms of Service

We hate legalese, so we've tried to make our Terms of Service readable. If you've got any questions, feel free to [ask us](mailto:support@growstuff.org), and we'll do our best to answer.

ai
0
0
Growstuff
GUARDRAILS.md

📰 AI News Daily — 09 Dec 2025

- Google unveils Gemini-powered AI glasses launching in 2026, signaling a major wearable comeback.

aiagentllm
0
0
inai-sandy
GUARDRAILS.md

AI Chatbot Integration Guide

This guide covers the AI-powered conversational features in Wolfbot, including context-aware chat, memory management, and safety features.

airagprompt
0
0
Jorak01
GUARDRAILS.md

GangGPT - AI-Powered GTA V Multiplayer Server

GangGPT is a revolutionary Grand Theft Auto V multiplayer server that transforms traditional roleplay gaming through advanced artificial intelligence integration. Built on the RAGE:MP framework with Azure OpenAI GPT-4o-mini, this project creates a living, breathing virtual world where every interaction is enhanced by intelligent systems.

airagopenai
0
0
dragoscv
GUARDRAILS.md

Twitter/X Launch Thread

**Timing:** Post entire thread Wednesday morning (24h after HN)

aillmprompt
0
0
base76-research-lab
GUARDRAILS.md

Implementing AI-Safety in a LLM-System Architecture

title: Implementing AI-Safety in a LLM-System Architecture

aillmrag
0
0
marcpre
GUARDRAILS.md

DeepSeek R1: Case Study in Failed Extrinsic Alignment

**Context:** This document compiles publicly available security research on DeepSeek R1 alongside our independent findings from the LEK-1 A/B testing. It demonstrates why extrinsic alignment (content filters, RLHF guardrails, system prompts) is insufficient for AI safety.

aiprompteval
0
7
Snider
Page 1 of 3