Hallucination detection 3

Detects factual correctness of content, especially AI-generated content. Bullshit Detector detects if the content is factually correct. It is designed to identify AI-generated content that may contain factually incorrect information, also known as 'hallucinations'. The tool aims to help users distinguish between accurate and misleading information produced by AI models.
AI simulation environment for testing LLM apps at scale. Snowglobe is a simulation environment for LLM teams designed to test how their AI applications respond to real-world user behavior. It enables users to run full workflows through realistic scenarios, catch edge cases early, and confidently improve model performance before deploying to production. Snowglobe helps AI teams test LLM apps at scale by simulating real-world conversations, uncovering risks, and improving overall model performance.
AI platform for battle-testing and improving AI agents. Janus is an advanced AI platform designed to battle-test and improve AI agents. It conducts thousands of AI simulations against chat and voice agents to surface critical failures such as hallucinations (fabricated content), rule violations (policy breaches), and tool-call/performance failures. Janus offers custom evaluations, personalized datasets, and actionable insights to help users detect and mitigate risky agent behavior, ensuring model reliability and performance.