Data as of Jul 25, 2026 · Based on 2,718,867 AI responses across 9,511 prompts · See how Parse measures this
NVIDIA/FasterTransformer is a repository providing highly optimized transformer-based encoder and decoder components for inference, built on CUDA and supporting models like BERT, GPT, and T5. Development has transitioned to TensorRT-LLM, and the repository will remain available without further updates.
Parse Score