devtool
AI/ML benchmark for local LLM inference and XGBoost training on GPU/CPU
This is a very basic and unpolished collection of benchmarking scripts, lacking documentation and unique value, unsuitable for serious performance evaluation.
3.3Overall
Utility4
Onboarding3
Craft4
Niche fit3
Longevity2
Five dimensions scored independently; overall is a weighted average. Scores are only comparable within this same rubric.
Good for
Developers who need a quick, informal personal project to roughly test GPU/CPU performance.
Not for
Anyone or any team requiring reliable, repeatable, and well-documented benchmark results.
Project description
github.com
Alternatives
MLPerfOpen LLM Leaderboard (for LLM evaluation)custom scripts with better documentation and methodology