Data as of Jul 25, 2026 · Based on 2,718,867 AI responses across 9,511 prompts · See how Parse measures this
LLMart is a toolkit for evaluating large language model robustness through adversarial testing, built with PyTorch and Hugging Face integrations to enable scalable red teaming attacks. It supports configurable attack patterns, soft prompt optimization, and parallelized optimization across multiple devices for both high-level evaluation and research experimentation.
Parse Score