ai-ml · Found in 5 repositories
nemo-evaluator
Evaluates LLMs across 100+ benchmarks from 18+ harnesses (MMLU, HumanEval, GSM8K, safety, VLM) with multi-backend execution. Use when needing scalable evaluation on local Docker, Slurm HPC, or cloud p
View source: zechenzhangAGI/AI-research-SKILLs ↗Install from source
Install using Skill Manager:
sk install https://github.com/zechenzhangAGI/AI-research-SKILLs/tree/main/11-evaluation/nemo-evaluator/SKILL.mdSource-path status is inferred from metadata; it does not verify a live download. Scan and quality scores describe registry checks and are not a guarantee of safety.
Attribution and permissions
Check the original repository for permission and license terms before reuse. Registry metadata does not grant a license.
This guide links to the author’s instructions and includes metadata only.
Request removal or correct attributionCopies with matching content
Exact Markdown body copies across 5 repositories. This count does not identify the original author.
- MesferAli/XCircle/.claude/skills/nemo-evaluator/SKILL.md
- Orchestra-Research/AI-Research-SKILLs/11-evaluation/nemo-evaluator/SKILL.md
- Orchestra-Research/AI-research-SKILLs/11-evaluation/nemo-evaluator/SKILL.md
- ihatesea69/HieuNghi-AI-Skills/airesearch_skills/11-evaluation/nemo-evaluator/SKILL.md
- zechenzhangAGI/AI-research-SKILLs/11-evaluation/nemo-evaluator
- zechenzhangAGI/AI-research-SKILLs/11-evaluation/nemo-evaluator/SKILL.md