lucataco/nsfw_video_detection
Detect NSFW content in videos. Accepts a video input and classifies it as 'nsfw' or 'normal' for automated content moder...
Found 19 models (showing 1-19)
Detect NSFW content in videos. Accepts a video input and classifies it as 'nsfw' or 'normal' for automated content moder...
Detect NSFW content in images, returning a binary label ('normal' or 'nsfw'). Accepts a single image as input and output...
Detects inappropriate content in images including nudity, explicit content, hentai, violence, and other NSFW categories....
Analyzes images and text prompts to detect content policy violations including public figures, copyright concerns, and N...
Determines the toxicity level of text-to-image prompts using a fine-tuned Llama-13b model. Analyzes input prompts and re...
Moderate text prompts, model responses, and images for safety compliance. Accepts text with optional multiple images and...
Classifies content in LLM prompts and responses for safety and harm detection. Built on Llama 3 with 8 billion parameter...
Classify text content for safety and content moderation based on the MLCommons hazards taxonomy. Built on Llama-3.1-8B a...
Detect hate speech and toxic content in text. Accepts a text string and returns JSON scores for toxicity, severe_toxicit...
Classify images for policy violations across sexually_explicit, dangerous_content, and violence_gore categories. Takes a...
Classify the safety of multimodal inputs (image and user message) for content moderation. Accepts an image (required) an...
Detect NSFW content in images. Accepts an image and outputs a safety label indicating whether the content is safe or uns...
Detect NSFW content in images. Takes an image and returns a moderation result with nsfw_detected (boolean), a list of NS...
Moderate text prompts and assistant responses for safety and policy compliance. Accepts a user message (prompt) and/or a...
Analyzes images for inappropriate content using MiniCPM-V-2.6 and returns structured safety assessments with detailed cl...
Detect NSFW content in images and compare results across two classifiers. Takes an image input and returns JSON safety f...
Classify text content based on custom safety policies you provide. Takes a text prompt and your written safety policy as...
Classifies text content based on custom safety policies written in plain English. Trained specifically for safety reason...
Classifies text content for safety and moderation, analyzing user prompts and assistant responses to determine if they a...