Open-source multimodal AI provenance detection platform with explainable evidence cards, provider consensus, and benchmark-driven quality gates.
-
Updated
Sep 7, 2026 - Python
Open-source multimodal AI provenance detection platform with explainable evidence cards, provider consensus, and benchmark-driven quality gates.
Public profile README covering how I approach product, AI systems, and customer signal in production.
Open-source behavioral intelligence for child safety moderation.
Crash course for new tirreno developers. Open-source security framework architecture, integration guide, and risk rules for developers and product teams.
MCP server with 41 AI-powered tools for child safety, fraud detection, and content moderation. Detects bullying, grooming, sextortion, romance scams, social engineering, and more. For Claude, Cursor, and MCP-compatible AI assistants.
Official .NET SDK for Moderyo Content Moderation API
ARGUS — open-source trust & safety platform. Fake account detection, deepfake identification, and disinformation monitoring for digital advertising.
GenAI governance and AI agent launch-readiness framework with golden evals, risk gates, automation boundaries, human escalation, and local readiness reports.
Autonomous Trust & Safety system using multi-agent orchestration on AWS. Detects, investigates, and enforces policy violations at scale with human-in-the-loop for sensitive cases. Built with Lambda, Step Functions, Bedrock, DynamoDB, and React.
Give your AI agents the power to discover, coordinate, and collaborate safely. Open-core collaboration layer for AI agents.
Synthetic trust and safety operations case study for AI user reports, covering risk taxonomy, severity classification, SQL analysis, dashboarding, QA, escalation workflow, and responsible automation.
Official Moderyo Unity SDK — AI-powered content moderation for games. UPM package for Unity 2021.3+.
A curated list of tools, research, organizations, and frameworks for trust and safety, child protection, and content moderation.
Technical Lead case study: React Native iOS app, trust & safety, tutoring marketplace, App Store deployment
Free and open-source dating platform that respects your privacy
A documented product exploration around small events, trust signals and continuity after people meet.
Raid AI — independent third-party profile of a public API surface, by API Evangelist. Raid AI (RAID AI, Inc.) is an enterprise-grade deepfake and AI-content forensics platform covering audio, image, document, and video. Its REST API detects AI-generated, cloned, deepfaked, and digitally edited media — returning a verdict, a confidence score, and (o
To associate your repository with the trust-safety topic, visit your repo's landing page and select "manage topics."