all publishers

qhjqhj00

@qhjqhj00 source repo

7,582 published skills · page 76 of 76

  1. ▌
    Pixel Art · qhjqhj00 bundle
    Generates pixel art SVG illustrations for READMEs, docs, or slides, with character templates, chat bubbles, and scene composition guidance.
    3 repo stars
  2. ▌
    Pennylane · qhjqhj00 bundle
    Train quantum circuits with automatic differentiation and build hybrid quantum-classical models using PennyLane, including VQE, QAOA, and integration with PyTorch, JAX, and TensorFlow.
    3 repo stars
  3. ▌
    Geopandas · qhjqhj00 bundle
    Performs geospatial vector data analysis with GeoPandas, including reading/writing shapefiles, GeoJSON, GeoPackage, and PostGIS, geometric operations, spatial joins, overlays, coordinate transformations, and map visualization.
    3 repo stars
  4. ▌
    Auroc · qhjqhj00
    Computes the AUROC metric using torchmetrics, handling binary, multiclass, and multilabel tasks with configurable thresholds and averaging.
    3 repo stars
  5. ▌
    Cider · qhjqhj00
    Computes CIDEr and related metrics to score how well generated image descriptions align with human consensus, using reference sentences and triplet annotations.
    3 repo stars
  6. ▌
    Menli · qhjqhj00
    Evaluates the robustness and alignment with human judgment of reference-based and reference-free evaluation metrics for machine translation and summarization, particularly under adversarial conditions.
    3 repo stars
  7. ▌
    Polos · qhjqhj00
    Scores generated image captions against reference captions and source images using the Polos metric, which is trained to align with human judgments and probes hallucination robustness and open-vocabulary evaluation.
    3 repo stars
  8. ▌
    Score · qhjqhj00
    Audits medical LLM benchmarks across five lifecycle phases using 46 medically tailored criteria to assess clinical relevance, data integrity, safety-critical capabilities, validity, and governance.
    3 repo stars
  9. ▌
    Spice · qhjqhj00
    Evaluates image captions by converting them into scene graphs and computing an F-score over semantic propositions, measuring how well a generated caption captures the meaning of an image compared to human references.
    3 repo stars
  10. ▌
    Squad · qhjqhj00
    Computes the SQuAD metric using torchmetrics, given predictions and ground truth. Use when evaluating question-answering outputs with exact match and F1 scores.
    3 repo stars
  11. ▌
    Ttsds · qhjqhj00
    Evaluates text-to-speech systems by measuring distributional distance between synthetic and real speech across five factors, producing a scalar score without subjective MOS ratings.
    3 repo stars
  12. ▌
    Visor · qhjqhj00
    Evaluates text-to-image models on spatial relationship accuracy using the VISOR metric, separating object detection from spatial correctness to reveal biases like object priority and merging.
    3 repo stars
  13. ▌
    Umap Learn · qhjqhj00 bundle
    Reduce high-dimensional data with UMAP for visualization, clustering preprocessing, and supervised or semi-supervised learning, including parameter tuning guidance.
    3 repo stars
  14. ▌
    Distributed LLM Pretraining Torchtitan · qhjqhj00 bundle
    Pretrains large language models at scale using PyTorch-native torchtitan with 4D parallelism, Float8, and distributed checkpointing.
    3 repo stars
  15. ▌
    Matplotlib · qhjqhj00 bundle
    Create publication-quality static, animated, and interactive plots with fine-grained control over every element, from basic charts to multi-panel figures, with export to PNG, PDF, and SVG.
    3 repo stars
  16. ▌
    125 Yuv Design · qhjqhj00 bundle
    Applies a battle-tested bilingual web-design system with typography, responsive rules, and performance patterns for Yuval Avidani's projects.
    3 repo stars
  17. ▌
    Bleurt · qhjqhj00
    Evaluates the correlation between automatic text generation scores and human quality ratings, including robustness to domain and quality drift, using metrics like Kendall's Tau and Pearson correlation.
    3 repo stars
  18. ▌
    Infolm · qhjqhj00
    Computes the InfoLM metric from torchmetrics for evaluating text generation against ground truth, with configurable information measures and sentence-level scoring.
    3 repo stars
  19. ▌
    L Eval · qhjqhj00
    Benchmarks long-context language models across 20 sub-tasks spanning 3k–200k tokens, covering retrieval, reasoning, summarization, and instruction understanding, with exact-match accuracy as the primary metric.
    3 repo stars
  20. ▌
    Lambre · qhjqhj00
    Scores generated text for morphosyntactic well-formedness by measuring how closely it adheres to language-specific dependency rules extracted from treebanks.
    3 repo stars
  21. ▌
    Logauc · qhjqhj00
    Computes the LogAUC metric using the torchmetrics implementation for binary, multiclass, or multilabel classification tasks.
    3 repo stars
  22. ▌
    Recall · qhjqhj00
    Computes the Recall metric using torchmetrics, including configuration for binary, multiclass, and multilabel tasks.
    3 repo stars
  23. ▌
    Reflex · qhjqhj00
    Evaluates machine-generated log summaries without human-written references, using LLM judgment and dense embeddings to score relevance, informativeness, and coherence.
    3 repo stars
  24. ▌
    Stream · qhjqhj00
    Evaluates spatial realism and temporal flow consistency of AI-generated videos using embedding spaces and Fourier transforms, producing bounded STREAM-S and STREAM-T scores.
    3 repo stars
  25. ▌
    Vpeval · qhjqhj00
    Evaluates text-to-image generation models by decomposing assessment into five specialized skills (object presence, count, spatial relations, scale, and text rendering) and open-ended prompts, producing interpretable binary scores with visual and textual explanations.
    3 repo stars
  26. ▌
    Paper Audit · qhjqhj00 bundle
    Provides reviewer-style audit and deep review for academic papers in LaTeX, Typst, and PDF formats, producing structured issue bundles, pass/fail gates, and revision roadmaps.
    3 repo stars
  27. ▌
    Self Review · qhjqhj00 bundle
    Reviews an academic paper using the NeurIPS review form with three reviewer personas, ensemble scoring, and reflection refinement. Extracts text from PDF, runs structured review, and outputs actionable feedback.
    3 repo stars
  28. ▌
    Tensorboard · qhjqhj00 bundle
    Visualize training metrics, debug models with histograms, compare experiments, visualize model graphs, and profile performance with TensorBoard.
    3 repo stars
  29. ▌
    A3 Eval · qhjqhj00
    Benchmarks mobile GUI agents on multi-step tasks across 20 Android apps, measuring task completion and essential-state navigation with Task Success Rate and Essential State Achieved Rate.
    3 repo stars
  30. ▌
    Apessrc · qhjqhj00
    Evaluates the faithfulness of abstractive summaries by verifying if factual claims (masked as cloze questions) in the reference summary can be correctly answered using only the generated summary, compared against a gold-standard answer derived from the source context.
    3 repo stars
  31. ▌
    Epsilon · qhjqhj00
    Evaluates the correlation between a zero-cost NAS metric (epsilon) and actual training accuracy across different neural architecture search spaces, testing the metric's ability to rank architectures without training. It probes whether output dispersion from constant weight initializations can serve as a reliable.
    3 repo stars
  32. ▌
    Latency · qhjqhj00
    Measures inference latency of binarized, 8-bit, and 32-bit convolutional layers on edge devices to evaluate the efficiency and speedup of the Larq Compute Engine framework compared to standard implementations.
    3 repo stars
  33. ▌
    Ndcg 10 · qhjqhj00
    Evaluates how well internal model representations (hidden states) predict token-level information importance in summarization tasks, using NDCG@10 and Spearman's rank correlation.
    3 repo stars
  34. ▌
    R2score · qhjqhj00
    Computes the R2Score metric using torchmetrics, handling single and multi-output predictions with options for adjusted and variance-weighted scores.
    3 repo stars
  35. ▌
    Runtime · qhjqhj00
    Benchmarks inference latency and computational runtime of transformer models and MLX operations across Apple Silicon and NVIDIA GPU backends, with configurable input lengths and batch sizes.
    3 repo stars
  36. ▌
    T5 Eval · qhjqhj00
    Benchmarks a text-to-text transformer across GLUE, SuperGLUE, CNN/Daily Mail, SQuAD, and WMT, reporting GLUE average, BLEU, ROUGE-2-F, and Exact Match scores.
    3 repo stars
  37. ▌
    Theilsu · qhjqhj00
    Computes Theil's U (uncertainty coefficient) between predictions and ground truth using the torchmetrics implementation, handling categorical data and NaN strategies.
    3 repo stars
  38. ▌
    Tpr Fpr · qhjqhj00
    Evaluates speaker verification models by computing true positive rate at fixed false positive rate thresholds, probing embedding space separation of same-speaker versus different-speaker pairs.
    3 repo stars
  39. ▌
    Usfiscaldata · qhjqhj00 bundle
    Query the U.S. Treasury Fiscal Data API for federal financial data including national debt, government spending, revenue, interest rates, exchange rates, and savings bonds. Access 54 datasets and 182 data tables with no API key required.
    3 repo stars
  40. ▌
    Copilot Docs · qhjqhj00 bundle
    Configure repository-specific guidance for GitHub Copilot by creating and structuring .github/copilot-instructions.md files.
    3 repo stars
  41. ▌
    Abc Eval · qhjqhj00
    Benchmarks large language models on symbolic music understanding and instruction following using text-based ABC notation, covering syntax parsing, error detection, segment-level reasoning, and sequence-level musical analysis.
    3 repo stars
  42. ▌
    Accuracy · qhjqhj00
    Evaluates an AI judge system's pairwise ranking accuracy on generated commit messages against a heuristic ground truth from five automatic text metrics, using the MCMD dataset.
    3 repo stars
  43. ▌
    Adp Eval · qhjqhj00
    Benchmarks LLM agents fine-tuned with the Agent Data Protocol across software engineering, web browsing, OS/database tool use, and reasoning tasks, reporting unit test pass rates and task success rates.
    3 repo stars
  44. ▌
    Anderson · qhjqhj00
    Computes the Anderson-Darling test statistic and p-value using scipy.stats.anderson for evaluating predictions against ground truth.
    3 repo stars
  45. ▌
    Ape Eval · qhjqhj00
    Benchmarks automatic post-editing (APE) models on WMT'18 SMT, SubEdits, and MLQE-PE datasets, reporting BLEU, ChrF, and TER scores computed with SacreBLEU and TERCOM.
    3 repo stars
  46. ▌
    Arc Eval · qhjqhj00
    Benchmarks systems on the Abstraction and Reasoning Corpus (ARC) by requiring inference of abstract transformation rules from few input-output grid demonstrations and application to novel test cases, reporting the fraction of tasks solved.
    3 repo stars
  47. ▌
    Art Eval · qhjqhj00
    Benchmarks medical AI agents on synthetic EHR tasks, measuring success rates for data retrieval, temporal aggregation, and threshold-based conditional logic with exact-match scoring.
    3 repo stars
  48. ▌
    Jq · qhjqhj00 bundle
    Query, filter, transform, and aggregate JSON data using jq, with practical patterns for shell pipelines and integration with CLI tools.
    3 repo stars
  49. ▌
    Hqq Quantization · qhjqhj00 bundle
    Quantize LLMs to 8/4/3/2/1-bit precision without calibration data, using multiple backends and HuggingFace/vLLM integration.
    3 repo stars
  50. ▌
    Adhx · qhjqhj00 bundle
    Fetches any X/Twitter post as structured JSON via the ADHX API, including full article content, author info, and engagement metrics, without scraping or a browser.
    3 repo stars
  51. ▌
    Tmux · qhjqhj00 bundle
    Manage persistent terminal sessions, windows, and panes with tmux, including scripting multi-pane layouts and automating commands from bash.
    3 repo stars
  52. ▌
    Aeon · qhjqhj00 bundle
    Performs time series machine learning tasks including classification, regression, clustering, forecasting, anomaly detection, segmentation, and similarity search using the aeon toolkit.
    3 repo stars
  53. ▌
    Shap · qhjqhj00 bundle
    Explains machine learning model predictions using SHAP values, covering feature importance, visualization plots, model debugging, bias analysis, and production deployment.
    3 repo stars
  54. ▌
    PPTX · qhjqhj00 bundle
    Creates, edits, and analyzes PowerPoint presentations by converting to markdown, unpacking raw XML, and applying design principles for new slides.
    3 repo stars
  55. ▌
    Qutip · qhjqhj00 bundle
    Simulate open and closed quantum systems with QuTiP, covering master equations, Lindblad dynamics, decoherence, and quantum optics.
    3 repo stars
  56. ▌
    Pymoo · qhjqhj00 bundle
    Solve single- and multi-objective optimization problems with NSGA-II/III, MOEA/D, and other evolutionary algorithms, including Pareto front analysis, constraint handling, and benchmarking on standard test problems.
    3 repo stars
  57. ▌
    Mlflow · qhjqhj00 bundle
    Track ML experiments, manage the model registry with versioning, deploy models, and reproduce experiments using MLflow's framework-agnostic platform.
    3 repo stars
  58. ▌
    Review · qhjqhj00 bundle
    Routes quality reviews to specialized critic agents based on file type or flags, covering peer review, code review, and manuscript polish.
    3 repo stars
  59. ▌
    Depmap · qhjqhj00 bundle
    Query the Cancer Dependency Map (DepMap) for CRISPR gene dependency scores, drug sensitivity data, and gene effect profiles to identify cancer-specific vulnerabilities, synthetic lethal interactions, and validate oncology drug targets.
    3 repo stars
  60. ▌
    Hf MCP · qhjqhj00 bundle
    Connects AI assistants to the Hugging Face Hub via MCP server tools to search models, datasets, Spaces, and papers, retrieve repository details and documentation, run compute jobs, and use Gradio Spaces as AI tools.
    3 repo stars
  61. ▌
    Slidev · qhjqhj00 bundle
    Create and present web-based slidedecks for developers using Slidev with Markdown, Vue components, code highlighting, animations, and interactive features.
    3 repo stars
  62. ▌
    Plotly · qhjqhj00 bundle
    Creates interactive Plotly visualizations in Python, covering Express and Graph Objects for scatter, line, bar, heatmap, 3D, and geographic charts, plus subplots, styling, and HTML export.
    3 repo stars
  63. ▌
    Seaborn · qhjqhj00 bundle
    Create publication-quality statistical graphics in Python with dataset-oriented plotting, semantic mapping, and built-in statistical estimation.
    3 repo stars
  64. ▌
    Diagram Skills · qhjqhj00 bundle
    Provides guides for creating diagrams and visualizations using Mermaid, Excalidraw, PlantUML, TikZ, and other tools, covering flowcharts, architecture diagrams, and scientific illustrations.
    3 repo stars
  65. ▌
    Pydicom · qhjqhj00 bundle
    Read, write, and manipulate DICOM medical imaging files, including pixel data extraction, metadata editing, anonymization, format conversion, and compression handling.
    3 repo stars
  66. ▌
    Wiki QA · qhjqhj00 bundle
    Answers questions about a code repository by analyzing source files and citing evidence with linked citations.
    3 repo stars
  67. ▌
    Phoenix Observability · qhjqhj00 bundle
    Self-hosted observability platform for LLM applications, providing tracing, evaluation, datasets, experiments, and real-time monitoring to debug and improve AI systems.
    3 repo stars
  68. ▌
    Auc · qhjqhj00
    Evaluates machine learning classifiers on their ability to distinguish signal from background in particle physics simulations, measuring how well algorithms rank signal events above background ones using the AUC metric.
    3 repo stars
  69. ▌
    Eas · qhjqhj00
    Validates the Emotional Attitude Score (EAS) metric by measuring its consistency with human judgment on word-level sentiment polarity, using the AmbGIMT dataset and pairwise score comparisons.
    3 repo stars
  70. ▌
    Eer · qhjqhj00
    Compute the Equal Error Rate (EER) metric using torchmetrics for binary, multiclass, or multilabel classification tasks, with reference signatures and usage examples.
    3 repo stars
  71. ▌
    Fid · qhjqhj00
    Measures distributional similarity between original GAN-generated images and their semantically manipulated counterparts using the Fréchet Inception Distance (FID) metric.
    3 repo stars
  72. ▌
    Mos · qhjqhj00
    Evaluates the naturalness, speaker similarity, and real-time synthesis speed of a Mandarin speech cloning system across diverse practical application scenarios.
    3 repo stars
  73. ▌
    Qqe · qhjqhj00
    Computes bibliometric indices for AI/NLP conferences, including QQE, average/median citations, and citation inequality, from annual publication and citation data.
    3 repo stars
  74. ▌
    Roc · qhjqhj00
    Computes the Receiver Operating Characteristic (ROC) metric using torchmetrics, supporting binary, multiclass, and multilabel tasks.
    3 repo stars
  75. ▌
    Sdr · qhjqhj00
    Quantifies audio source separation quality by computing the signal-to-distortion ratio (SDR) between ground-truth and estimated stems, with per-stem and record-level averaging.
    3 repo stars
  76. ▌
    Tec · qhjqhj00
    Measures the trade-off between computation time and energy consumption in mobile edge computing by computing a weighted sum of the two objectives, given system configuration parameters and per-user task characteristics.
    3 repo stars
  77. ▌
    Ara Compiler · qhjqhj00 bundle
    Compiles any research input — PDF papers, GitHub repositories, experiment logs, code directories, or raw notes — into a complete Agent-Native Research Artifact (ARA) with cognitive layer (claims, concepts, heuristics), physical layer (configs, code stubs), exploration graph, and grounded evidence. Use when ingesting a.
    3 repo stars
  78. ▌
    Ray Data · qhjqhj00 bundle
    Process large ML datasets in parallel across CPU or GPU clusters, with streaming execution, multi-format I/O, and integration with Ray Train, PyTorch, and TensorFlow for batch inference and preprocessing pipelines.
    3 repo stars
  79. ▌
    Psnr · qhjqhj00
    Evaluates the trade-off between file size reduction and image fidelity when encoding radio astronomy data using JPEG2000, benchmarking both lossless and lossy compression modes to determine the compression ratio at which visual artifacts first appear.
    3 repo stars
  80. ▌
    Flops · qhjqhj00
    Evaluates computational throughput and real-time efficiency of embedded CPU and GPU platforms by measuring peak FLOPS via a matrix rotation kernel and assessing inference latency and power consumption on a robotic vision pipeline.
    3 repo stars
  81. ▌
    F1score · qhjqhj00
    Compute the F1Score metric using torchmetrics when predictions and ground-truth labels are available.
    3 repo stars
  82. ▌
    Kruskal · qhjqhj00
    Compute the Kruskal-Wallis H-test using scipy.stats.kruskal for independent samples, returning the H statistic and p-value.
    3 repo stars