IndicGuard
IndicGuard is a multilingual content safety guardrail model for Indic languages, built as a LoRA adapter on top...
IndicGuard is a multilingual content safety guardrail model for Indic languages, built as a LoRA adapter on top...
A small, fast content-moderation classifier fine-tuned from Qwen2.5-3B-Instruct for Discord-style chat. It is the L1 fast-triage tier of...
HiveTraceGuard-Pro is a compact Russian-first guardrail built on Qwen3-0.6B for fast input and output classification. Built for LLMs...
Ориентированная на выравнивание тонкая настройка DeepSeek-R1-Distill-Qwen-1.5B. Улучшение выравнивания измеряется прозрачно с использованием фреймворка Bloom от Anthropic. Три из...
A lightweight security guard model fine-tuned from Qwen3-4B for detecting prompt injections, enforcing AI agent guardrails, and identifying...
Helion-V1-Reasoning is a conversational Reasoning AI model designed to be helpful, harmless, and honest. The model focuses on...
Helion-V1 is a conversational AI model designed to be helpful, harmless, and honest. The model focuses on providing...
Introducing the GA Guard series — a family of open-weight moderation models built to help developers and organizations...
Это модель создания подсказки для взлома на основе текстов очков знаний. Модель обучена на наборе данных Llama-2-7b и...
Генеративная модель, обученная классифицировать подсказки по различным категориям безопасности и генерировать практические правила. Архитектура модели: T5ForConditionalGeneration. Данные: prosocial-dialog...