#
open-llm-leaderboard
Here are 3 public repositories matching this topic...
RetardBench is an open, no-censorship benchmark that ranks large language models purely on how retarded they are.
jailbreak large-language-models llm prompt-injection red-teaming-tools ollama llm-evaluation uncensored-llm open-llm-leaderboard llm-jailbreaks prompt-injection-llm-security ai-benchmark ai-red-teaming llm-benchmark
-
Updated
Mar 2, 2026 - TypeScript
EnsembleX utilizes the Knapsack algorithm to optimize Large Language Model (LLM) ensembles for quality-cost trade-offs, offering tailored suggestions across various domains through a Streamlit dashboard visualization.
python benchmark knapsack huggingface streamlit large-language-models llm llm-evaluation open-llm-leaderboard
-
Updated
May 5, 2024 - Python
Add this topic to your repo
To associate your repository with the open-llm-leaderboard topic, visit your repo's landing page and select "manage topics."