Model Training & Fine-tuning
-
qhjqhj00 Skill FlopsEvaluates computational throughput and real-time efficiency of embedded CPU and GPU platforms by measuring peak FLOPS via a matrix rotation kernel and assessing inference latency and power consumption on a robotic vision pipeline.
Audited 3 -
qhjqhj00 Skill SquadComputes the SQuAD metric using torchmetrics, given predictions and ground truth. Use when evaluating question-answering outputs with exact match and F1 scores.
Audited 3 -
qhjqhj00 Skill LambreScores generated text for morphosyntactic well-formedness by measuring how closely it adheres to language-specific dependency rules extracted from treebanks.
Audited 3 -
qhjqhj00 Skill VpevalEvaluates text-to-image generation models by decomposing assessment into five specialized skills (object presence, count, spatial relations, scale, and text rendering) and open-ended prompts, producing interpretable binary scores with visual and textual explanations.
Audited 3 -
qhjqhj00 Skill ApessrcEvaluates the faithfulness of abstractive summaries by verifying if factual claims (masked as cloze questions) in the reference summary can be correctly answered using only the generated summary, compared against a gold-standard answer derived from the source context.
Audited 3 -
qhjqhj00 Skill Ndcg 10Evaluates how well internal model representations (hidden states) predict token-level information importance in summarization tasks, using NDCG@10 and Spearman's rank correlation.
Audited 3 -
qhjqhj00 Skill Tpr FprEvaluates speaker verification models by computing true positive rate at fixed false positive rate thresholds, probing embedding space separation of same-speaker versus different-speaker pairs.
Audited 3 -
qhjqhj00 Skill Adp EvalBenchmarks LLM agents fine-tuned with the Agent Data Protocol across software engineering, web browsing, OS/database tool use, and reasoning tasks, reporting unit test pass rates and task success rates.
Audited 3 -
qhjqhj00 Skill Bss EvalEvaluates speech language models on beyond-semantic speech attributes such as dialect comprehension, multi-turn context memory, emotion perception, age-aware response generation, and non-verbal cue handling, reporting accuracy and judge-based scores.
Audited 3 -
ichichuang Bundle Pytorch FsdpProvides expert guidance on PyTorch Fully Sharded Data Parallel (FSDP) training, covering parameter sharding, mixed precision, CPU offloading, and FSDP2.
Audited 0 -
smith6jt-cop Skill Pytorch Common PitfallsFixes common PyTorch bugs including percentile calculations, LayerNorm for Conv1d, and buffer edge cases in reinforcement learning and neural network code.
Audited 3 -
nvidia Bundle Tao Train Pose ClassificationTrain, evaluate, export, and run inference for pose classification models using ST-GCN on skeleton keypoint sequences.
Audited 2.2k -
nvidia Bundle Deepstream Profile PipelineProfile a DeepStream pipeline with Nsight Systems and derive its configs from the measurement.
2.2k -
nvidia Bundle Earth2studio Create DiagnosticCreate Earth2Studio diagnostic model wrappers for single-step data transformations, including simple derived diagnostics, packaged AutoModel diagnostics, and generative or diffusion diagnostics.
2.2k -
orchestra-research Bundle DeepspeedProvides expert guidance for distributed training with DeepSpeed, covering ZeRO optimization stages, pipeline parallelism, FP16/BF16/FP8, 1-bit Adam, and sparse attention.
10.4k -
orchestra-research Bundle Ml Training RecipesProvides battle-tested PyTorch training recipes for LLMs, vision, diffusion, and biomedical domains, covering training loops, optimizer selection, LR scheduling, mixed precision, and debugging.
Audited 10.4k -
orchestra-research Bundle Model PruningCompress large language models by 40-60% with minimal accuracy loss using one-shot pruning techniques like Wanda and SparseGPT, enabling faster inference and deployment on constrained hardware.
10.4k -
orchestra-research Bundle Evaluating Code ModelsEvaluates code generation models across HumanEval, MBPP, MultiPL-E, and 15+ benchmarks with pass@k metrics. Use when benchmarking code models, comparing coding abilities, testing multi-language support, or measuring code generation quality.
10.4k -
orchestra-research Bundle Knowledge DistillationCompress large language models using knowledge distillation from teacher to student models, covering temperature scaling, soft targets, reverse KLD, logit distillation, and MiniLLM training strategies.
Audited 10.4k -
google-gemma Bundle Gemma TrainerFine-tune Gemma models locally using QLoRA, Unsloth, or TRL for SFT, DPO, and reward modeling, with dataset preparation and conversion to GGUF or LiteRT-LM.
-
majiayu000 Bundle DpoTrains language models with Direct Preference Optimization using preference pairs, covering DPOTrainer setup, dataset preparation, and beta tuning for stable preference learning without explicit reward models.
Audited 567 -
ecnu-icalk Skill Ciou GiouReplaces GIoU with Complete IoU (CIoU) loss in PyTorch object tracking or detection tasks, combining overlap area, center-point distance, and aspect-ratio similarity for improved bounding-box regression.
Audited 559 -
lingxling Bundle ShapExplains machine learning model predictions using SHAP values, covering feature importance, visualizations, debugging, bias analysis, and production deployment.
Audited 253 -
lingxling Bundle GtarsHigh-performance toolkit for genomic interval analysis in Rust with Python bindings. Use when working with genomic regions, BED files, coverage tracks, overlap detection, tokenization for ML models, or fragment analysis in computational genomics and machine learning applications.
Audited 253 -
lingxling Bundle QutipSimulate and analyze quantum mechanical systems, including open quantum systems, using QuTiP's solvers for master equations, Lindblad dynamics, and quantum trajectories.
Audited 253 -
lingxling Bundle DatamolPythonic wrapper around RDKit for cheminformatics, simplifying SMILES parsing, standardization, descriptors, fingerprints, clustering, 3D conformers, and parallel processing while returning native rdkit.Chem.Mol objects.
253 -
lingxling Bundle MolfeatConvert chemical structures (SMILES or RDKit molecules) into numerical representations for machine learning, covering 100+ featurizers including ECFP, MACCS, descriptors, and pretrained models like ChemBERTa, with support for QSAR modeling and virtual screening.
Audited 253 -
lingxling Skill Hf MCPConnects AI assistants to the Hugging Face Hub via MCP server tools to search models, datasets, Spaces, and papers, retrieve repo details and documentation, run compute jobs, and use Gradio Spaces as AI tools.
253 -
nimoqup046-collab Skill AI MlOrchestrates AI/ML development workflows covering LLM applications, RAG systems, AI agents, ML pipelines, and observability.
Audited 2 -
lord1egypt Skill Slime Rl TrainingGuides LLM post-training with RL using slime, a Megatron+SGLang framework for training GLM, Qwen, DeepSeek, and Llama models with GRPO, async, and multi-turn workflows.
Audited 2 -
scoheart Bundle LLM Price LookupLook up and compare LLM API pricing across Models.dev, OpenRouter, and official provider pages, including token costs and aliases.
2 -
neuralblitz Skill PytorchProvides guidance on using PyTorch for deep learning, covering tensors, autograd, nn.Module, DataLoaders, and best practices.
Audited 1 -
cloudthinker-ai Skill Managing RayManages Ray clusters, jobs, Serve deployments, and distributed workloads via the Dashboard API and CLI, with discovery-first checks and safety guardrails.
7 -
tools-only Bundle Time Series RegressionAeon provides time series regressors across 9 categories for predicting continuous values from temporal sequences.
Audited 7 -
infinition Bundle Llama CppRun GGUF models locally with llama.cpp, including finding the right file on the Hugging Face Hub, installing, quantizing, serving, and using Python bindings.
2 -
phoroth Skill LangfuseInstrument LLM applications with Langfuse for tracing, prompt versioning, evaluation, and dataset management across Python and JavaScript SDKs.
3