MindTrial

by petmal

MindTrial: Evaluate and compare AI language models (LLMs) on text-based tasks with optional file/image attachments and tool use. Supports multiple providers (OpenAI, Google, Anthropic, DeepSeek, Mistral AI, xAI, Alibaba, Moonshot AI, OpenRouter), custom tasks in YAML, and HTML/CSV reports.

โ˜… 5 stars1 forksActiveGo
73
Good

๐Ÿ“Š Score Breakdown

๐Ÿ›ก๏ธSecurity30%
5.0/5
โšกUtility30%
1.0/5
๐Ÿ”„Maintenance25%
5.0/5
๐Ÿ’ŽUniqueness15%
4.0/5

Overall = Security (30%) + Utility (30%) + Maintenance (25%) + Uniqueness (15%). Full methodology โ†’

โ„น๏ธ Details

Ecosystem

Claude Skill

Language

Go

Pricing

Free

License

MPL-2.0

Status

Active

๐Ÿ“ˆ GitHub Signals

5

Stars

1

Forks

0

Commits (30d)

0

Open Issues

Last commit: 5 months ago

ai-benchmarkai-evaluation-toolsai-model-comparisonai-toolanthropicartificial-intelligence-projectsdeepseekgoogle-gemini-aigrok-ailanguage-models-aillm-benchmarkingllm-comparisonllm-evaluation-frameworkmistral-aimoonshot-aiopenaiopenrouteropensourceqwenxai

๐Ÿ… Show your score

Scored 73/100 for security, utility and maintenance. Add the badge to your README or site to show it, verified by an independent directory.

MindTrial scored 73/100 on SkillsIndex
Markdown
[![MindTrial scored 73/100 on SkillsIndex](https://skillsindex.dev/api/badge/petmal-mindtrial)](https://skillsindex.dev/tools/petmal-mindtrial/)
HTML
<a href="https://skillsindex.dev/tools/petmal-mindtrial/"><img src="https://skillsindex.dev/api/badge/petmal-mindtrial" alt="MindTrial scored 73/100 on SkillsIndex" height="20"></a>

Know before you install ๐Ÿ“ฌ

We score every tool 0-100 for security, maintenance and utility. Get the weekly shortlist of the highest-scored, vetted tools, plus an alert when a package you rely on goes stale. Free.

Data last verified: 5 months ago. See something wrong? Report it โ†’