What is mle-bench?
mle-bench is a AI tool that MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering. It has a Nerq Trust Score of 71/100 (B). 1.3K GitHub stars. Published by Unknown. Last analyzed August 2026.
Why This Score
- ⚠️ Security: 0/100 — Some security concerns
- ⚠️ Maintenance: 0/100 — Maintenance activity is low
- ✅ Community: 1.3K stars, 0 downloads — Large community
- ⚠️ Transparency: License: Not specified — No license specified
Trust & Safety Overview
What mle-bench Does
mle-bench is a agent in the AI tool category. MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering. It is published by an independent developer and has no specified license. With 1.3K GitHub stars and 0 downloads, it has a growing community of users and contributors.
Who Should Use mle-bench
mle-bench is well-suited for production use given its strong trust score and active community.
Details
| Author | Unknown |
|---|---|
| Category | AI tool |
| License | Not specified |
| Type | agent |
| Source | View on GitHub |
| Security Score | 0/100 |
| Activity Score | 0/100 |
How to Get Started
Check the trust score before installing:
curl nerq.ai/v1/preflight?target=openai-mle-bench
Setup guide · Full safety report · Production review · Is it safe?
Safer Alternatives
| Tool | Trust | Stars |
|---|---|---|
| openclaw | 59 | 218.2K |
| stable-diffusion-webui | 62 | 160.7K |
| prompts.chat | 73 | 145.8K |
| generative-ai-for-beginners | 66 | 106.7K |
| ComfyUI | 69 | 103.7K |
Frequently Asked Questions
Last updated August 2026. Trust scores based on automated analysis of public data.