Results for “conviction-scoring”
51 skillsMore results
batch-vs-realtime-scoring
The choice is not about scale or sophistication.
2
lead-scoring
Score and prioritize leads based on firmographic fit and behavioral engagement signals, producing ranked tiers for sales team focus. Use when the user requests lead scoring or provides relevant inputs for this workflow.
159
eval-judge
Score LLM and agent outputs using LLM-as-judge techniques — direct scoring against rubrics or pairwise comparison between two outputs. Includes built-in bias mitigation for position bias, length bias, and self-enhancement bias. Load when the user asks to score an output, judge a response, evaluate against a rubric, compare two outputs, do direct scoring, run pairwise comparison, or says "rate this", "which response is better", "score this against the rubric", "judge this output", "LLM as judge this". Sub-skill of eval-output orchestrator.
3 · bundle
lead-scoring
Defines ideal customer profile filters, scores inbound and outbound leads, and builds a lightweight qualification rubric to sharpen pipeline focus for founder-led sales.
20
cx-qa-appeal-process
Use to design or audit a QA dispute and appeal workflow with timeboxes, adjudication standards, and second-level consistency so appeals improve trust instead of rewriting scores without rules. Trigger for "QA appeal process", "agents disputing scores", "who adjudicates QA disputes", "overturn rate too high", second-level review standards, or calibration erosion from ad-hoc score changes.
1
cumprimento-de-sentenca
Promove o cumprimento de sentenca (titulo executivo judicial) — disposicoes gerais (CPC 513), titulos judiciais (CPC 515), cumprimento DEFINITIVO de pagar quantia com intimacao em 15 dias e multa de 10% + honorarios de 10% se nao pago (CPC 523 §1, penhora/avaliacao §3), demonstrativo atualizado do debito (CPC 524), cumprimento PROVISORIO sob recurso sem efeito suspensivo (CPC 520-522), alimentos com prisao civil (CPC 528 §3) e Fazenda Publica (CPC 534-535). Use quando o operador disser cumprimento de sentenca, executar a sentenca, intimar para pagar em 15 dias, multa de 10% cumprimento, cobrar sentenca transitada.
6
connotation-cop
Police the project's vocabulary — bust vague terms, keep the CONTEXT.md glossary sharp, and lock in decisions worth remembering as ADRs. Use when the user debates naming, says "what should we call this", asks to pin down terminology, wants a decision recorded, or when another skill (hot-seat, whiteboard) surfaces a decision that clears the ADR bar. Just reading the glossary for vocabulary is NOT this skill — trigger only when the words or decisions are being changed.
0 · bundle
self-eval
Honestly evaluate AI work quality using a two-axis scoring system with mandatory devil's advocate reasoning and cross-session anti-inflation detection.
20.4k
lead-qualifier
Multi-dimensional lead qualification scoring. Evaluates leads against BANT criteria, firmographic fit, behavioral signals, and intent indicators. Outputs qualified/disqualified verdict with detailed reasoning.
2 · bundle
benchmark-methodology
Scores competitors across nine weighted dimensions with explicit 1–5 rubrics and a tension plot, producing comparable profile cards for competitive analysis.
226k
f1score
Compute the F1Score metric using torchmetrics when predictions and ground-truth labels are available.
3
draft-score
Lightweight ContentShake AI self-check the /draft stage can call before saving. Returns just SEO + Quality scores (no full optimization) so the writer knows whether the draft is in winning territory before /quality-check runs. Fails soft when SEMRUSH_API_KEY is unset.
0
jnov-motion
Drafts a Motion for Judgment Notwithstanding the Verdict (JNOV) under FRCP 50(b) or state equivalents, with alternative new-trial request. Builds element-by-element evidentiary insufficiency arguments using transcript citations and preserves the appellate record. Use when drafting JNOV motions, post-trial motions, renewed judgment as a matter of law, or challenging jury verdicts for insufficient evidence.
34
bison-strategy
BISON v2.0 — Conviction Holder (Hardened). Top 10 assets by volume. All signals are score contributors — no hard gates. Scanner enters via create_position internally (Wolverine pattern). RatchetStop exits. Thesis exit REMOVED. v2.0: every hard gate converted to score contributor, ensureExecutionAsTaker=false, conviction-scaled margin 25-37%.
1 · bundle
lead-scorer
Score raw leads as HOT/WARM/COOL based on config-driven weights from agency.config.json
2 · bundle
ai-writing-detector
Score any piece of writing for AI-generation tells and produce a weighted 0-100 scorecard with flagged evidence and ranked fixes. Use this skill whenever the user asks "does this sound AI-written", "run this through the AI detector", "score this writing", "check this for AI tells", "would this pass as human", "humanize check", or wants any article, blog post, email, or copy audited for AI patterns before publishing. Also use it when the user pastes or points to text and asks how it reads, whether it's too "ChatGPT-ish", or wants a QA pass on generated content. Works on pasted text, files, and URLs.
0 · bundle
strike-zone-analyst
A funnel and account-scoring diagnostic engine for any sales org. Connect a CRM and a product-analytics tool (plus optional enrichment, community, and meeting tools). Three modes. (1) FUNNEL DIAGNOSIS finds leaky conversion gates by channel with per-stage leakage, dollarized leverage points, and cohort velocity. (2) SPRINT PLANNING enriches qualified accounts into a ranked backlog with verified buying committees. (3) SCORING AUDIT finds where your scoring model is missing real ICPs. Trigger on 'funnel diagnosis', 'diagnose the funnel', 'where are we leaking', 'why is [channel] underperforming', 'conversion by channel', 'sprint planning', 'score these accounts', 'find missed ICPs', 'audit the scoring model', or any channel-level cohort-conversion, account-prioritization, or scoring-gap question.
0 · bundle
accuracy
Evaluates an AI judge system's pairwise ranking accuracy on generated commit messages against a heuristic ground truth from five automatic text metrics, using the MCMD dataset.
3
arbitrage-scanner
Detect and evaluate arbitrage opportunities across sportsbooks and prediction markets. Calculate guaranteed-profit scenarios, middle bets, and cross-platform price discrepancies. Use when comparing odds across books, calculating arb percentages, evaluating middle opportunities, or building odds-comparison workflows. Also trigger for 'arb bet', 'sure bet', 'arbitrage', 'odds comparison', 'middle bet', 'risk-free bet', or 'line shopping'.
0
impugnacao-ao-cumprimento
Redige a impugnacao ao cumprimento de sentenca (CPC 525) — defesa do executado em titulo JUDICIAL, no prazo de 15 dias apos os 15 do art. 523 e independentemente de penhora, com as hipoteses do §1 (I-VII: nulidade de citacao, ilegitimidade, inexigibilidade, penhora/avaliacao, excesso de execucao, incompetencia, causa extintiva superveniente), dever de declarar o valor correto sob pena de rejeicao (§4-5) e efeito suspensivo so com garantia + grave dano (§6). Use quando o operador disser impugnar cumprimento, impugnacao ao cumprimento de sentenca, excesso de execucao, defender de cumprimento de sentenca, executado quer se defender de sentenca.
6
eval
Evaluate and rank agent results by metric or LLM judge for an AgentHub session. Use when the user runs /hub:eval or asks to score, compare, or pick a winner among completed AgentHub agents.
11
verdict-form
Drafts civil trial verdict forms with sequentially numbered jury questions covering liability, affirmative defenses, damages, comparative fault, and special interrogatories. Enforces plain-language phrasing, logical conditional flow, and jurisdiction-appropriate formatting. Use when preparing verdict forms, special verdict forms, jury interrogatories, or general verdict forms with interrogatories.
34
aikido-triage
Aikido Findings Triage Workflow
21
cognovit-note
Drafts cognovit promissory notes with confession of judgment provisions, gated by mandatory jurisdictional enforceability research, usury compliance, and statutory disclosure requirements. Advises on alternatives where cognovit clauses are prohibited. Use when drafting cognovit notes, confession of judgment instruments, or loan documents requiring waiver-of-defense provisions.
34
predictions
Use when making a forward-looking claim with a checkable outcome (reply within 24h, error rate will drop, this skill will see more use) — record to state/predictions.jsonl with a review horizon so reflection can grade you later. Closes the in-the-moment double-loop.
6 · bundle
cx-incentive-design
Use to design support incentives that improve behaviour without destroying the metric — pairing pay with guardrails, naming gaming modes, and choosing measures that survive Goodhart pressure. Trigger for "incentive plan", "agent bonus scheme", "SPIFF design", "pay for QA score", "what metric should we bonus", CSAT incentives, or reviewing whether a comp change is driving gaming.
1
prioritizing-vulnerabilities-with-cvss-scoring
Calculate CVSS scores, interpret vector strings, and prioritize vulnerabilities using CVSS alongside EPSS and CISA KEV for effective risk-based remediation.
24.6k · bundle
recall
Computes the Recall metric using torchmetrics, including configuration for binary, multiclass, and multilabel tasks.
3
vendor-risk-scoring
Vendor privacy risk tiering methodology for processor management. Covers scoring factors including data volume, sensitivity, transfer locations, certifications, breach history, and control maturity with weighted risk calculation and tier assignment.
228 · bundle
eval
Evaluate and rank agent results by metric or LLM judge for an AgentHub session. Use when the user runs /hub:eval or asks to score, compare, or pick a winner among completed AgentHub agents.
2
conan-skill
《名侦探柯南》式少年推理叙事:线索陈列、假设排除、指证演说节奏、正义框架; 剧本与桌游推理主持参考,虚构,禁止可模仿犯罪细节、禁止冒充刑侦。 触发:柯南、推理、排除法、真相只有一个(结构层,勿抄录版权长对白)等。
9 · bundle
harness-rehearse
Rehearse
18 · bundle
c2c-eval
Benchmarks language model agents on the C2C multi-agent negotiation task, reporting win rate across starting positions.
3
alphagbm-marks-cycle
Provides a single 0-100 cycle score blending VIX, SPY IV Rank, Put/Call ratio, and valuation percentile to determine offense vs. defense posture, based on Howard Marks' market cycle framework.
1.2k
stanley-druckenmiller-investment
Synthesizes outputs from 8 upstream market analysis skills into a unified conviction score (0-100), pattern classification, and allocation recommendation for Druckenmiller-style portfolio positioning.
2.3k · bundle