Browser Tool Tests LLM Security Using Deterministic Grading Without an AI Judge
A developer has released The AI Crash Test, a browser-based tool that evaluates large language models against adversarial prompts using deterministic, code-based grading rather than a secondary AI judge. Users supply their own API key, which is sent directly to the model provider and never passes through the tool's server — a claim the developer says anyone can verify via browser DevTools in about 30 seconds. The grading engine, called gradecore, produces byte-identical scores across repeated runs, eliminating variability from temperature or judge drift. The tool covers eight tasks across seven attack categories, including prompt injection, hallucination baiting, and refusal calibration, generating a severity-weighted vulnerability report. The developer acknowledges that more comprehensive red-teaming tools such as NVIDIA's garak and Microsoft's PyRIT exist, positioning this project not as a new category but as a narrower, auditable alternative focused on reproducibility.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in