huggingface-community-evals

huggingface/huggingface-community-evals · Agent Skill (multi-file)

by Hugging Face · bundle

Published · Last updated


Run evaluations for Hugging Face Hub models using inspect-ai and lighteval on local hardware, with backend selection between vLLM, Transformers, and accelerate.

SKILL.md

Files

This skill is a package of 6 files. Install with the command above, or download the folder.

  • 📄SKILL.md entry
  • 📁examples
  • 📄.env.example GitHub 0 B
  • 📄USAGE_EXAMPLES.md 2.0 KB
  • 📁scripts
  • ⚙️inspect_eval_uv.py 2.9 KB
  • ⚙️inspect_vllm_uv.py 8.9 KB
  • ⚙️lighteval_vllm_uv.py 9.0 KB

Related

  1. awq-quantization · orchestra-research bundle
    Quantize large language models to 4-bit using activation-aware weight quantization, achieving ~3x speedup with minimal accuracy loss for deployment on limited GPU memory.
    10.4k
    repo stars
  2. awq-quantization · majiayu000 bundle
    Quantize large language models to 4-bit precision using activation-aware weight quantization, reducing memory footprint and speeding up inference with minimal accuracy loss.
    567
    repo stars
  3. llamaguard · orchestra-research
    Deploy Meta's LlamaGuard moderation model to filter LLM inputs and outputs across 6 safety categories using HuggingFace, vLLM, or FastAPI.
    10.4k
    repo stars
  4. transformers · k-dense-ai bundle
    Load pre-trained models from Hugging Face Hub, run pipeline inference, generate text, and fine-tune models on NLP, vision, audio, and multimodal tasks using the Transformers library.
    30.2k
    repo stars
  5. fine-tuning-expert · jeffallan bundle
    Fine-tune LLMs using LoRA, QLoRA, and PEFT with Hugging Face, including dataset preparation, hyperparameter tuning, evaluation, and deployment.
    10.4k
    repo stars
  6. tao-finetune-huggingface-model · nvidia bundle
    Fine-tune HuggingFace CV, VLM, or LLM models on local NVIDIA GPUs using an NGC PyTorch container, with support for full or LoRA training, dataset handling, and optional model push to the Hub.
    2.2k
    repo stars

Frequently asked questions

How do I install the huggingface-community-evals skill?

Run npx skillmds add huggingface/huggingface-community-evals in your terminal (requires Node.js), paste this page's agent-chat prompt into Claude, Cursor, or any MCP-connected agent, or download the SKILL.md file and copy it into your agent's skills directory.

What does the huggingface-community-evals skill do?

Run evaluations for Hugging Face Hub models using inspect-ai and lighteval on local hardware, with backend selection between vLLM, Transformers, and accelerate. It is listed under AI & ML, Model Training & Fine-tuning on SkillMD.

Is huggingface-community-evals safe to use?

SkillMD's automated safety review verdict for this skill is CAUTION. Independent scanners report: SkillSpector: PASS, Skill Scanner: PASS. Capability flags: executes scripts, reads secrets. SkillMD never runs a skill's scripts for you; review the SKILL.md before installing.

Which AI agents work with huggingface-community-evals?

This skill is tagged as working with Claude Code, Claude.ai, OpenAI Codex. SKILL.md is an open format, so most agents that read a skills directory can load it too.

Is huggingface-community-evals free to use?

Yes. Installing skills from SkillMD is free, and the skill stays under its author's original license.

Who published huggingface-community-evals?

Hugging Face (@huggingface) published this skill as a verified publisher. Their other Agent Skills are listed on their SkillMD profile.