Skip to content
#

llm-comparison

Here are 54 public repositories matching this topic...

MindTrial: Evaluate and compare AI language models (LLMs) on text-based tasks with optional file/image attachments and tool use. Supports multiple providers (OpenAI, Google, Anthropic, DeepSeek, Mistral AI, xAI, Alibaba, Moonshot AI, OpenRouter), custom tasks in YAML, and HTML/CSV/JSON reports.

  • Updated Sep 6, 2026
  • Go

Benchmark abierto en español de 170 modelos de IA (118 con 20+ runs, 69 rankeados, juez Phi-4 independiente). Calidad, costo, velocidad, long-context y fuga de credenciales como dimensiones separadas. Alternativas a Claude, GPT y Gemini para agentes n8n/Hermes. Calculadora interactiva con tus propios pesos.

  • Updated Sep 14, 2026
  • Python

Add this topic to your repo

To associate your repository with the llm-comparison topic, visit your repo's landing page and select "manage topics."

Learn more