discovered 03 Aug 2026
malicious-gpt
→ View on GitHubMalla is a research tool designed to analyze and evaluate 220 malicious large language model (LLM) applications, such as WormGPT and FraudGPT. Its primary use case is to provide insights into the performance, authorship attribution, and security vulnerabilities of these LLMs through various claims and datasets, including quality assessments and jailbreak prompt evaluations. Notable features include a comprehensive dataset of malicious prompts and responses, as well as tools for benchmarking LLM responses against potential security threats.