HolobiomicsLab
- 7.4k skills
- 0 followers
- 19 hours ago last updated
- ▌ Eic Signal Peak Detection 2 · holobiomicslabUse when after EIC candidate generation from LC/HRMS data (mzXML, mzML, or netCDF formats), when you need to localize discrete peaks within chromatographic profiles and assign retention time boundaries, apex intensities, and quality scores prior to peak annotation or cross-sample alignment.
- ▌ Metadata Field Validation 2 · holobiomicslabUse when you have received new or updated MassBank records (in plain-text or structured format) that must be integrated into the MassBank-data repository and you need to ensure they conform to the MassBank format specification before acceptance.
- ▌ Dimple Pipeline Execution 2 · holobiomicslabUse when when you have deposited mass spectrometry imaging data in NetCDF (CDF) format paired with MATLAB workspace files (.
- ▌ Image Intensity Jittering 2 · holobiomicslabUse when when preparing ion images (single-channel 2D arrays or multi-channel spectral images) from mass spectrometry imaging for contrastive learning in DeepION's COL or ISO modes.
- ▌ Spatial Feature Embedding 2 · holobiomicslabUse when when analyzing imaging mass spectrometry datasets where you need to reduce high-dimensional peak intensity features while preserving spatial structure, and when automatic peak picking and marker ion identification are required.
- ▌ Lipid Identification Scoring 2 · holobiomicslabUse when after peak picking has generated a peaklist of experimental fragment m/z values (from Q-Exactive orbitrap, Agilent/Bruker/SCIEX Q-TOF UHPLC-HRMS/MS, or direct infusion/imaging experiments) and you need to compare those fragments against the LipidMatch in-silico fragmentation library.
- ▌ M Z Alignment Across Samples 2 · holobiomicslabUse when after mass track construction for individual samples, when you need to establish consensus m/z values across a cohort of LC-MS samples to build a unified feature table.
- ▌ Ms Ms Spectrum Preprocessing 2 · holobiomicslabUse when you have raw MS/MS spectra (in MGF or mzML format) with unscaled peak intensities and noise artifacts, and you plan to rank chemical formulas, predict adducts, or score precursor–spectrum agreement using a machine learning model such as MIST-CF.
- ▌ Candidate Metabolite Ranking 2 · holobiomicslabUse when when you have generated a set of predicted metabolite structures from BioTransformer's metabolism prediction engine and need to assign identity to observed compounds from LC-MS/MS, spectral, or chromatographic experiments.
- ▌ Cohort Performance Reporting 2 · holobiomicslabUse when you have NMR metabolite measurements from peripheral blood samples (plasma/serum) paired with processing delay metadata (pre-centrifugation and post-centrifugation times) and need to benchmark metabolic parameter stability across delay windows.
- ▌ Command Line Tool Invocation 2 · holobiomicslabUse when you need to bootstrap a tool workflow by generating a version- or instrument-specific default configuration file (e.g., for MS-DIAL 4 vs. 5), execute an analysis on formatted input files (e.g., MS-DIAL export .txt files), or capture tool output for downstream validation.
- ▌ Dataset Integrity Assessment 2 · holobiomicslabUse when when you have downloaded a released version of a structured dataset (e.g., LOTUS from Zenodo) and need to confirm it matches the documented headline statistics before downstream analysis, or when auditing data integrity after ingestion into a processing pipeline.
- ▌ Exact Mass Database Matching 2 · holobiomicslabUse when after feature detection and alignment on raw MS data, when you have a list of unknown feature m/z values and need to assign them to known xenobiotic metabolites or their predicted biotransformation products.
- ▌ Gnps Library Format Assembly 2 · holobiomicslabUse when you have extracted MS1 and MS2 scans (in mzML/mzXML format) from raw chromatogram files and possess user-provided metadata (retention time, m/z, compound name, molecular weight, annotation fields) that must be combined into a single structured library entry suitable for spectral library.
- ▌ Hybrid Model Fusion Strategy 2 · holobiomicslabUse when you have 1H NMR spectral data from complex mixtures and need to identify component compounds, but a single architecture (CNN or Transformer alone) fails to capture both fine local patterns in peak structures and long-range dependencies across the full spectral range.
- ▌ Linux Command Line Execution 2 · holobiomicslabUse when you have vendor-specific raw mass spectrometry data (ThermoFisher .raw, Agilent .
- ▌ Lipid A Structure Annotation 2 · holobiomicslabUse when when you have high-resolution tandem mass spectrometry (MS2) data in .ms2 format and need to identify and annotate lipid A structures at scale.
- ▌ M Z Alignment Across Samples 3 · holobiomicslabUse when you have extracted mass tracks (EICs) from multiple LC-MS samples at 0.001 amu resolution and need to construct a sample-agnostic m/z reference frame.
- ▌ Mass Feature To Node Mapping 2 · holobiomicslabUse when you have an untargeted metabolomics feature table with m/z values, retention times, and intensity measurements, a metabolic network representation with compound nodes and chemical formulas, and you want to infer functional pathway activity directly from features without performing.
- ▌ Metabolite Spectral Matching 2 · holobiomicslabUse when you have an experimental mass spectrum (or a set of spectra from LC-MS/MS data) and need to identify the underlying metabolite(s) by comparing against known reference spectra in GNPS or a local indexed repository.
- ▌ Ms Ms Spectrum Preprocessing 3 · holobiomicslabUse when you have raw or semi-processed MS/MS spectral data from bottom-up tandem mass spectrometry experiments (data-dependent acquisition) that you intend to input to de novo peptide sequencing tools like Casanovo.
- ▌ Network Node Label Spreading 2 · holobiomicslabUse when you have an untargeted metabolomics dataset with a two-layer network topology already constructed (one layer representing biochemical knowledge/pathways, the other representing data-driven MS2 similarity), seed metabolites with reliable annotations from database matching or curation, and.
- ▌ Pathway Database Integration 2 · holobiomicslabUse when you have intensity measurements (peak features, protein intensities, or gene expression values) with compound or gene annotations (KEGG IDs, ChEBI IDs, UniProt IDs, or ENSEMBL IDs), and you need to aggregate them into biologically meaningful pathway groups for differential analysis.
- ▌ Precursor Product Mz Parsing 2 · holobiomicslabUse when you have raw MRM sample files from a LC-MS/MS instrument and need to systematically recover all precursor m/z and product m/z pairs for each transition. Use this as an initial parsing step before quantitation, method optimization, or transition verification workflows.
- ▌ Qc Sample Quality Assessment 2 · holobiomicslabUse when after drift correction and before imputation when you have LC-MS data with designated QC samples and you need to remove features with poor reproducibility across QC replicates.
- ▌ R Package Function Execution 2 · holobiomicslabUse when you have raw Bruker NMR spectral data files (1D 1H format) stored in a directory structure and need to prepare them for automated metabolite identification and quantification in ASICS.
- ▌ Retention Time Peak Matching 2 · holobiomicslabUse when after drift correction and quality flagging, when you have a feature abundance matrix with associated metadata (Feature_ID, m/z, retention time) and need to identify which features likely represent the same underlying metabolite or adduct series before statistical analysis or metabolite.
- ▌ Simulation Output Validation 2 · holobiomicslabUse when after executing a multi-stage simulation workflow in R and/or MATLAB, when you have generated intermediate and final outputs (tables, figures, model objects) and need to confirm that all artifacts conform to the documented format, structure, and expected content before downstream analysis.
- ▌ Spectral Mz Window Filtering 2 · holobiomicslabUse when when you have resolved mzML or mzXML spectrum files and need to isolate signals for a target m/z value (e.g., 870.954) across all retention times or a specific scan.
- ▌ Sterol Isomer Classification 2 · holobiomicslabUse when you have LC-IM-MS/MS raw data from sterol-containing tissue samples and need to assign detected peaks to specific structural isomers (e.g., distinct double bond positions or saturation patterns in C27–C29 sterols).
- ▌ Stocsy Metabolite Assignment 2 · holobiomicslabUse when use STOCSY when you have preprocessed 1H NMR spectral data with an unidentified peak of interest (driver signal at a specific δ ppm value) and need to determine its metabolite identity by finding correlated signals across the spectrum.
- ▌ Two Layer Topology Traversal 2 · holobiomicslabUse when you have an untargeted metabolomics dataset with partial metabolite annotations (from database matching or prior curation) and need to extend annotation coverage to unannotated metabolites.
- ▌ YAML JSON Structural Parsing 2 · holobiomicslabUse when you have a versioned workflow definition file (YAML or JSON) from a specific release commit and need to verify it conforms to the project's schema specification, validate the presence of all required metadata fields (name, version, inputs, outputs, steps), and detect syntax errors or.
- ▌ Expert Review Preparation 2 · holobiomicslabUse when when a paper describes a computational or statistical method and you need to verify that claims are supported by available code, data, or documentation before human expert evaluation.
- ▌
- ▌ Spectral Metadata Grouping 2 · holobiomicslabUse when processing a mass spectrometry dataset (in FragHub JSON format or similar) where duplicate spectral records are suspected or known to exist. The input dataset should already be in a standardized format with computed or retrievable SPLASH keys.
- ▌ Validator Tool Integration 2 · holobiomicslabUse when you have a repository of structured records (e.g., mass spectrometry data, metadata, or domain-specific formats) and need to enforce validation rules systematically across all records.
- ▌ Python Environment Pinning 2 · holobiomicslabUse when when you have access to a research repository or README documenting a machine learning implementation (e.g., Keras/TensorFlow-based deep learning model) and need to reproduce the computational environment exactly. Triggers include: (1) README explicitly lists pinned versions (e.
- ▌ Nps Classification Prediction 2 · holobiomicslabUse when you have acquired a mass spectrum from an unknown suspected illicit drug analyte and need to compare it against a synthetic NPS database to rank candidate identities by similarity.
- ▌ Adduct Mass Offset Assignment 2 · holobiomicslabUse when when you have an LC-MS feature table with m/z and retention time columns and need to identify which observed ions correspond to the same neutral compound under different ionization conditions and isotopic enrichment.
- ▌ Adduct Regex Pattern Matching 2 · holobiomicslabUse when ingesting mass spectrometry spectra from heterogeneous databases or libraries where adduct annotations may be incomplete, incorrectly formatted, or inconsistent with the ionization mode. Use it before downstream analysis (e.
- ▌ Annotation Confidence Scoring 2 · holobiomicslabUse when after recursive annotation propagation has assigned metabolite labels to previously unannotated nodes in a two-layer metabolomic network, and before reporting final annotated metabolite identities.
- ▌ Binary Additive Flag Encoding 2 · holobiomicslabUse when constructing HPLC column feature vectors from raw metadata that includes additive composition flags (e.g., presence/absence or concentration of formic acid, acetic acid, TFA, or phosphoric acid in mobile phase eluents A and B).
- ▌ Cosine Similarity Computation 2 · holobiomicslabUse when when comparing two MS/MS spectra (query and reference) to quantify their spectral resemblance for compound identification or molecular networking, particularly when you need a simple, symmetric measure that is insensitive to precursor mass differences and does not require peak alignment.
- ▌ Cross View Similarity Scoring 2 · holobiomicslabUse when you have an experimental mass spectrum (query) and a set of molecular candidate structures, and you need to rank the candidates by how well their predicted spectral features match the query spectrum.
- ▌ Database Metadata Enumeration 2 · holobiomicslabUse when when you have downloaded a curated structure-organism dataset (such as LOTUS) and need to verify the reported counts of unique entities (source databases, organisms, structures, and their pairs) to confirm dataset integrity, assess data coverage, or reproduce published statistics in a.
- ▌ Dda Acquisition Data Handling 2 · holobiomicslabUse when you have raw or processed LC-MS/MS data from DDA mode acquisitions and need to extract, annotate, and structure MS/MS spectra with purity labels (or quality indicators) to serve as input to the DNMS2Purifier customized model training workflow, or to prepare data for purification of.
- ▌ Decision Tree Path Extraction 2 · holobiomicslabUse when you have a trained shallow decision tree on ChemEcho feature vectors (sparse, high-dimensional representations of tandem mass spectra peaks and neutral losses) and need to convert it into an interpretable, deployable query for a domain-specific language like MassQL.
- ▌ Deep Learning Model Inference 2 · holobiomicslabUse when you have preprocessed mass spectrometry spectra (tokenized m/z and intensity pairs or feature matrices) and a trained deep learning model checkpoint, and you need to classify unknown compounds or generate prediction confidence scores for structural novelty analysis.
- ▌ Eic Data Extraction From Xcms 2 · holobiomicslabUse when after running XCMS getEIC() to generate xcmsEIC objects and fillPeaks() to produce a filled xcmsSet object, before computing the 12 peak-quality metrics (Apex Max-Boundary Ratio, Elution Shift, FWHM2Base, Jaggedness, Modality, Symmetry, Sharpness, Gaussian Similarity, Retention-Time.
- ▌ Feature Identifier Assignment 2 · holobiomicslabUse when after constructing MetaboSet objects from Excel-formatted LC-MS peak tables and before drift correction or quality flagging.
- ▌ Flat File Parsing And Loading 2 · holobiomicslabUse when when you have published LOTUS flat files (TSV or compressed TSV.GZ) containing structure-organism pairs and need to enumerate unique structures, group by organism prevalence, or validate record counts against gold-standard benchmarks.
- ▌ Fragment Ion Mass Calibration 2 · holobiomicslabUse when when comparing experimental spectra to reference library spectra and fragment ion m/z values show systematic drift or measurement noise that could distort neutral loss peaks or cosine similarity scores.
- ▌ Gradient Performance Encoding 2 · holobiomicslabUse when when you have extracted retention times from the top detected MS1 features in a LC-MS run and need to evaluate whether the gradient spreads those compounds efficiently across the available chromatographic time window—particularly during iterative gradient optimization where you need a.
- ▌ In Silico Fragment Prediction 2 · holobiomicslabUse when you have a collection of compound structures in SDF format (e.g., DNA adduct structures) and need to systematically generate predicted fragment spectra across a defined ionization level and mass range to populate a reference spectral database or validate experimental fragmentation patterns.
- ▌ Ion Image Augmentation Design 2 · holobiomicslabUse when when preparing ion image data from mass spectrometry imaging for contrastive self-supervised representation learning, and you need to generate augmented image pairs that reflect either co-localization relationships between different molecular ions (COL mode) or isotopic relationships.
- ▌ Lc Ms Feature Quality Scoring 2 · holobiomicslabUse when immediately after peak detection and feature table generation from LC-MS data, when you need to rank or filter features by confidence before annotation or statistical analysis.
- ▌ Metabolite Annotation Scoring 2 · holobiomicslabUse when you have a feature table with candidate metabolite annotations (m/z, retention time, chemical identifiers) from MS/MS spectra or external tools (SIRIUS, GNPS), sample metadata linking samples to organisms, and you need to prioritize candidates by both annotation quality AND biological.
- ▌ Metaboset Object Manipulation 2 · holobiomicslabUse when when you have read LC-MS peak table data from Excel (or equivalent) into R and need to organize it into a structured object that tracks feature abundances, sample information (injection order, QC status), and feature metadata (mass, retention time, Feature_ID) simultaneously.
- ▌ Molecular Formula Calculation 2 · holobiomicslabUse when you have user-specified lipid class constraints (e.
- ▌ Ms2 Annotation Interpretation 2 · holobiomicslabUse when after GNPS spectral library search has returned matched chemical annotations (with m/z values and cosine similarity scores) for MS/MS spectra.
- ▌ Multi Platform Ms Integration 2 · holobiomicslabUse when you have untargeted metabolomics data from multiple MS instruments (e.
- ▌ Mzml File Parsing And Loading 2 · holobiomicslabUse when when you have centroided mzML format LC–MS files from multiple runs (e.
- ▌ Neural Network Model Training 2 · holobiomicslabUse when you have downloaded LC-MS spectral peak data (DOI 10.25345/C5FD2F or equivalent) and need to build a supervised deep neural network classifier to distinguish peak classes in mass spectrometry data.
- ▌ Nps Classification Prediction 3 · holobiomicslabUse when you have an unknown mass spectrum from a suspicious analyte and need to determine whether it matches a known NPS or a derivative thereof. The analyte's mass spectrum is available in MSP or equivalent format, and you have a core drug structure to enumerate derivatives from.
- ▌ Pre Trained Model Fine Tuning 2 · holobiomicslabUse when you have a small training dataset for molecular property prediction (e.g., <500 samples from PredRet or MoNA databases) and a pre-trained GNN model is available that was trained on a related, larger molecular corpus.
- ▌ Precursor M Z Based Filtering 2 · holobiomicslabUse when you have an unknown MS/MS query spectrum with a known or measured precursor m/z value and need to search a spectral library (local or public: GNPS, MASSBANK, DrugBANK) to annotate the compound.
- ▌ Qc Sample Type Classification 2 · holobiomicslabUse when when constructing a sample list from an Excel template for LC/GC-MS analysis, you must classify each QC sample by type before proceeding to plate layout and randomization steps.
- ▌ Quality Metrics Summarization 2 · holobiomicslabUse when after running QC analysis on NMR or MS metabolomic data and obtaining per-feature CV values, use this skill to validate that the dataset meets FDA thresholds (CV < 0.30 for discovery, CV < 0.15 for quantification) and to report the proportion of features meeting each threshold.
- ▌ Raw Ms Data Format Conversion 2 · holobiomicslabUse when you have raw UPLC-HRMS data from ThermoFisher or Agilent instruments and need to feed it into MSThunder for nontargeted pollutant identification. Your input is a vendor binary format (.raw or .d) that MSThunder cannot directly ingest. Environment constraints (e.
- ▌ Smiles Canonicalization Rdkit 2 · holobiomicslabUse when when processing raw SMILES strings from external databases or user input that may contain non-canonical tautomeric forms, variable stereochemical notation, or redundant representations of the same chemical structure.
- ▌ Smiles Parsing And Validation 2 · holobiomicslabUse when you have SMILES strings for candidate novel psychoactive substance structures and need to convert them into a machine-readable molecular representation before computing descriptors, generating mass spectra, or calculating chemical fingerprints.
- ▌ Spectral Embedding Generation 2 · holobiomicslabUse when you have a collection of pre-processed MS/MS spectra (binned, intensity-normalized) and a trained MS2DeepScore base network, and you need to compute structural similarity scores between spectrum pairs or visualize spectra in chemical space via dimensionality reduction (e.g., UMAP).
- ▌ Tensor Encoding Deep Learning 2 · holobiomicslabUse when when you have validated SMILES strings or RDKit molecule objects representing chemical structures and need to feed them into a pre-trained deep learning model (such as PS2MS, NEIMS, or DeepEI) that expects fixed-size numerical tensor inputs.
- ▌ C Python Interface Wrapping 2 · holobiomicslabUse when you have a mature C++ library (like OpenMS) with stable APIs that you want to make accessible from Python environments, and you need to preserve performance-critical C++ execution while supporting rapid prototyping or integration into Python-based data pipelines (e.
- ▌ Dna Adduct Characterization 2 · holobiomicslabUse when when you have a collection of DNA adduct compound structures in SDF format that requires validation for structural integrity and completeness, and you need to generate predicted fragment spectra at defined ionization levels and mass ranges for comparison against experimental mass.
- ▌ Local Maxima Identification 2 · holobiomicslabUse when you have raw LC-HRMS profile-mode data and need to identify candidate chromatographic peaks before classification or feature extraction.
- ▌ Openms API Surface Exposure 2 · holobiomicslabUse when when you need to make OpenMS C++ classes, functions, or data structures callable from Python code, or when verifying that a newly bound C++ component can be imported and instantiated without errors in a Python environment.
- ▌ Runtime Comparison Analysis 2 · holobiomicslabUse when when a new version or variant of a tool claims performance improvements over a prior version (e.g., MASST+ vs. MASST), and you need empirical evidence that the claimed speedup (e.g., ~100-fold reduction in search time) is real, reproducible, and quantifiable.
- ▌ Schema Conformance Checking 2 · holobiomicslabUse when you have a collection of records in a standardized format (e.g., MassBank plain-text or structured records) that must be validated before commit or publication.
- ▌ Semantic Metabolite Ranking 2 · holobiomicslabUse when you have an unknown metabolite with unknown mass spectrum and need to prioritize structural candidates from databases (PubChem, HMDB) by their likelihood of being the true compound.
- ▌ Contrastive Pair Generation 2 · holobiomicslabUse when you have raw ion images from MSI data and need to train a contrastive encoder to learn stable, mode-specific representations.
- ▌ Ppm Mass Accuracy Filtering 2 · holobiomicslabUse when when processing imzML/ibd Imaging Mass Spectrometry datasets and you need to extract ion density maps for specific analytes or isotopes. Apply this skill after importing the .imzML metadata and .
- ▌ Mass Spectrometry Data Parsing 2 · holobiomicslabUse when you have raw mass spectrometry files in standard formats (mzML, mzXML, msp, MGF, JSON) and need to extract precursor m/z values, fragment peaks, neutral losses, retention times, and compound metadata into a structured, queryable spectrum object representation before performing MS/MS.
- ▌ 4d Lcimmsms Feature Extraction 2 · holobiomicslabUse when you have raw LC-IM-MS/MS data files from sterol lipid analysis and need to identify unsaturated sterol isomers by matching experimental collision cross section values against a quantum chemistry calculation-assisted CCS prediction database.
- ▌ Artifact Checksum Verification 2 · holobiomicslabUse when when reproducing a prior software release (especially one generated by automated versioning tools like Semantic Release), you need to confirm that the artifacts produced in your environment match the original release byte-for-byte.
- ▌ Attention Mechanism Validation 2 · holobiomicslabUse when after instantiating a transformer encoder module for mass spectrometry data processing (e.g., in IDSL_MINT), before training on large MS/MS datasets or running inference on test spectra.
- ▌ Automated Lipid Identification 2 · holobiomicslabUse when you have high-resolution tandem mass spectrometry (MS2) data in .ms2 format and need to systematically identify and annotate lipid A molecular structures.
- ▌ Compound Ground Truth Matching 2 · holobiomicslabUse when when you have pre-computed embeddings for query and reference MS/MS spectra, computed their cosine similarity matrix, and need to measure retrieval success by verifying whether the correct compound (identified by SMILES string) appears in the top-1, top-5, or top-10 ranked candidates from.
- ▌ Cross Split Metric Aggregation 2 · holobiomicslabUse when when you have a pre-trained model and need to report stable, generalizable performance on a fixed training set with multiple held-out test splits. Specifically: when you have 10 (or n) random query/reference splits on the same dataset (e.
- ▌ Docker Container Orchestration 2 · holobiomicslabUse when your analysis requires msconvert or another ProteoWizard tool on macOS, but native installation is infeasible or licensing-restricted. You need to convert vendor raw mass spectrometry files (.raw) to the open mzML format without installing ProteoWizard directly on your system.
- ▌ Feature Group Adduct Detection 2 · holobiomicslabUse when you have a feature table from LC-MS analysis (containing m/z, retention time, and intensity values) and need to identify which detected features represent the same molecular species ionized under different adduction states.
- ▌ File Format Robustness Testing 2 · holobiomicslabUse when when processing MS spectral data from multiple open mass spectra libraries (OMSLs) in mixed formats (MSP, MGF, JSON, CSV), especially when source data exhibits missing fields, malformed entries, inconsistent adduct representations, or non-standard format variants that may cause silent.
- ▌ Frequency Distribution Binning 2 · holobiomicslabUse when you have loaded a table of entity–attribute pairs (e.
- ▌ Gaussian Peak Shape Evaluation 2 · holobiomicslabUse when after peak detection on a composite mass track has identified candidate peaks in a mass chromatogram, and before compiling the final feature table.
- ▌ HTTP Connectivity Verification 2 · holobiomicslabUse when you need to confirm that a documented web service URL is live and reachable before attempting to submit analysis jobs, download results, or integrate the service into an automated pipeline. Use it as a prerequisite check when the service documentation claims academic or public availability.
- ▌ Imputation Algorithm Selection 2 · holobiomicslabUse when you have a metabolomics dataset with left-censored missing values (e.g., below limit of quantification in LC/MS or GC/MS) and need to evaluate multiple imputation approaches.
- ▌ Interpretable Machine Learning 2 · holobiomicslabUse when when you have tandem mass spectra data and need to predict a binary molecular property (e.g., presence of a functional group like a sulfo group) while maintaining full interpretability of the model's decision logic.
- ▌ Mass Spectral Feature Grouping 2 · holobiomicslabUse when you have untargeted metabolomics MS/MS spectra from multiple features and need to identify which features belong to the same molecular family or are related by biotransformation.
- ▌ Mass Spectrometry Data Parsing 3 · holobiomicslabUse when you have received raw or vendor-converted centroid mzML files from LC-MS, GC-MS, or DI-MS platforms and need to extract MS1 spectra before building mass tracks, performing peak detection, or constructing composite feature maps.