← Tool Radar
devtool

AI/ML benchmark for local LLM inference and XGBoost training on GPU/CPU

This is a very basic and unpolished collection of benchmarking scripts, lacking documentation and unique value, unsuitable for serious performance evaluation.

3.3Overall
Utility4
Onboarding3
Craft4
Niche fit3
Longevity2

Five dimensions scored independently; overall is a weighted average. Scores are only comparable within this same rubric.

Good for

Developers who need a quick, informal personal project to roughly test GPU/CPU performance.

Not for

Anyone or any team requiring reliable, repeatable, and well-documented benchmark results.

Project description

github.com

Alternatives

MLPerfOpen LLM Leaderboard (for LLM evaluation)custom scripts with better documentation and methodology
Visit site1 · Stars at evalEvaluated 2026-05-16

Similar tools