Apps tagged with 'AI Safety'

All apps in Apps tagged with 'AI Safety' category. Use the filters below to narrow down your search. 
Copy a direct link to this comment to your clipboard
  1. Petri icon
     2 likes

    Petri is an alignment auditing agent for rapid, realistic hypothesis testing. It autonomously crafts environments, runs multi turn audits against a target model using human like messages and simulated tools, and then scores transcripts to surface concerning behavior.

    Cost / License

    • Free
    • Open Source (MIT)

    Platforms

    • Mac
    • Windows
    • Linux
    • Self-Hosted
    • US flagUnited States
    Petri screenshot 1
    Petri screenshot 1
    Petri screenshot 2
    +1
    Petri screenshot 3
  2. Wardstone icon
     Like

    Wardstone is an LLM firewall and AI guardrail API that protects AI applications from prompt attacks, harmful content, data leakage, and suspicious links in a single inference call with ~30ms latency.

    Cost / License

    • Freemium
    • Proprietary

    Platforms

    • Online
    • Software as a Service (SaaS)
    • GB flagUnited Kingdom
    Wardstone screenshot 1
    Wardstone screenshot 1
    Wardstone screenshot 2
  3. Guardamos icon
     Like

    Guardamos is a third-party content audit API for companies building AI-powered fitness, training, and wellness features.

    Cost / License

    • Paid
    • Open Source (MIT)

    Platforms

    • Self-Hosted
    Guardamos screenshot 1
  4. Privacy-focused browser extension monitoring conversations on top AI chat platforms, providing plain-English summaries and safety alerts instead of full transcripts.

    Cost / License

    • Freemium
    • Proprietary

    Application type

    Platforms

    • Online
    • Google Chrome
    • Software as a Service (SaaS)
    • GB flagUnited Kingdom
    The Halo Aware dashboard gives parents their family's week at a glance: how many AI sessions each child had, what they were broadly about, and anything worth a look, all as calm summaries rather than raw chat logs. It even suggests gentle conversation starters for later.
    Halo Aware's Insights view shows your child's AI usage patterns over time: how many conversations they've had, which platforms they use most, common topics, and shifts worth noticing, like a jump in usage or a change in what they're exploring. Understanding trends, not reading messages.
    When an AI conversation raises a genuine concern, Halo Aware's Alerts flag it clearly, showing what happened, which platform, and why it might matter, along with suggestions for what you can do. Serious moments surface to you without exposing every private chat.
    +2
    Halo Aware is never hidden from the child. The extension shows a clear status card on their browser: safety alerts are on, which platform is covered, and exactly what happens, the AI checks for safety concerns and a parent sees topic summaries, but the child's actual messages are not saved.
  5. Doberman icon
     Like

    Your AI's guard dog. Doberman sits at runtime, gating every input, output and tool call to stop unsafe or unintended actions before they execute.

    Cost / License

    Platforms

    • Self-Hosted
    • NZ flagNew Zealand
    Doberman screenshot 1
    Doberman screenshot 1
    Doberman screenshot 2
    +5
    Doberman screenshot 3
    5 alternatives
  6. ShieldGemma is a set of instruction tuned models for evaluating the safety of text and images against a set of defined safety policies. You can use this model as part of a larger implementation of a generative AI application to help evaluate and prevent generative AI...

    Cost / License

    • Free
    • Proprietary

    Platforms

    • Self-Hosted
    • Google Cloud Platform
    3 alternatives
  7. Aegisora icon
     Like

    Zero-trust runtime security & governance layer for autonomous AI agents — prompt injection firewall, PII masking, real-time policy enforcement.

    Cost / License

    • Freemium
    • Open Source (MIT)

    Platforms

    • Online
    • Software as a Service (SaaS)
    • DE flagGermany
    • European Union flagEU
    Operational Control for AI Agents: Enforce least-privilege tool calls and monitor runtime behavior seamlessly.
    Zero-Latency Proxy: Intercept and inspect tool requests on the fly without introducing execution bottlenecks inside your VPC.
    Granular Policy Enforcements: Intercept malicious tool actions, mask PII on the fly, and maintain strict security boundaries.
    +2
    Multi-Agent Fleet Governance: Scale autonomous workflows securely with centralized oversight and audit logs.