Results for “accuracy”
39 skillsExpo Skill Eval
Evaluates Expo skills end-to-end: trigger accuracy, generated code quality, and runtime screenshots on iOS simulator and Android emulator via Expo Go.
2.2k · bundle
Fasttext
编写评估FastText文本分类模型的Python函数,计算accuracy、F1、recall和precision指标,并处理特定格式的标签文本分割。
559
Data Quality Checker
Validates data quality with completeness, consistency, accuracy checks and generates quality reports
6 · bundle
Data Validate
QA a completed analysis for methodology, accuracy, and bias before it is shared, shipped, or acted on.
0
Coding Guide
Create medical coding guides for billing accuracy. TRIGGERS - Use when user needs help with coding-guide related tasks.
22
Coding Guide
Create medical coding guides for billing accuracy. TRIGGERS - Use when user needs help with coding-guide related tasks.
3
More results
Infographics
Creates data-driven infographics and charts as accessible SVG. Use when visualizing data, choosing a chart type, generating an SVG chart or infographic, or reviewing a visualization for clarity and accuracy.
0 · bundle
Caveman
Reduces token usage by ~75% by stripping filler words, articles, and pleasantries while preserving full technical accuracy. Activates on user commands like 'caveman mode' or 'be brief'.
20.4k · bundle
Data Quality Checker
Validate financial data quality in market analysis documents before publication, checking price scales, instrument notation, date accuracy, allocation totals, and unit usage.
2.3k · bundle
Design QA Checklist
Create systematic QA checklists to verify that design implementations match specifications across visual accuracy, layout, interaction, content, accessibility, and cross-platform categories.
1.7k
Caveman
Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler, articles, and pleasantries while keeping full technical accuracy. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman.
11 · bundle
Caveman
Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler, articles, and pleasantries while keeping full technical accuracy. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman.
16
Caveman
Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler, articles, and pleasantries while keeping full technical accuracy. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman.
2
Caveman
Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler, articles, and pleasantries while keeping full technical accuracy. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman.
5
Figma Implement Design
Translates Figma designs into production-ready application code with pixel-perfect accuracy, using the Figma MCP server to fetch design context, screenshots, and assets.
23.3k · bundle
Dimensional Analysis
Orchestrates a dimensional-analysis pipeline to annotate codebases with unit/dimension comments, discover dimensional vocabulary, and detect arithmetic bugs from unit mismatches or precision loss.
6k · bundle
Mpmath Python
Use for writing, reviewing, debugging, testing, or validating Python mpmath arbitrary-precision numerical code. Trigger on mpf, mpc, mp.dps, workdps, interval arithmetic, high-precision quadrature, root finding, special functions, matrices, inverse transforms, or precision/convergence failures. Do not use for ordinary NumPy vectorization, SymPy symbolic manipulation, decimal currency arithmetic, or machine-float code with no precision requirement.
0 · bundle
Alphago Deep Rl
Strategic patterns for solving intractable problems through cascading approximation, self-improvement, and heterogeneous evaluation from DeepMind's AlphaGo system
10 · bundle
Performance
Improves measured performance while preserving correctness, reliability, and maintainability.
0
Deep Interview
Socratic deep interview with mathematical ambiguity gating before explicit execution approval
1
Calculator
Performs arbitrary-precision arithmetic calculations including addition, subtraction, multiplication, division, and exponents. Use when the user asks to calculate, compute, or evaluate math expressions, or when precise decimal arithmetic is needed to avoid floating-point errors.
3 · bundle
Pragmatic Programmer
Apply meta-principles of software craftsmanship: DRY, orthogonality, tracer bullets, and design by contract to build systems that are easy to change, understand, and trust.
1.6k · bundle
Polish
Performs a final quality pass fixing alignment, spacing, consistency, and micro-detail issues before shipping. Use when the user mentions polish, finishing touches, pre-launch review, something looks off, or wants to go from good to great.
2
Spec Analysis
Perform a non-destructive cross-artifact consistency and quality analysis across spec.md, plan.md, and tasks.md. Identifying inconsistencies, duplications, ambiguities, and underspecified items.
2
Systematic Debugging
Diagnose failures from runtime evidence before editing code.
4
Clean Code
Pragmatic coding standards - concise, direct, no over-engineering, no unnecessary comments
3
Clean Code
Pragmatic coding standards - concise, direct, no over-engineering, no unnecessary comments
0
Llava Next Improved Reasoning Ocr And World Knowledge Arxiv
LLaVA-NeXT: Improved Reasoning, OCR, and World Knowledge
6
Caveman
Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler, articles, and pleasantries while keeping full technical accuracy. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman.
0
Icp Analyst
Turn "is this account a real fit?" into a stacked verdict. A composite 0-100 fit score built from every signal you have, a breakdown by source, override flags for when a scoring tool is wrong, a channel classification, and a written rationale. Built for reps, marketers, and RevOps, customizable to your CRM and whatever scoring you use. Trigger on "is {account} a real fit?", "score this list against our profile", "find lookalikes to our best customers", "why is the score wrong on {account}?", "validate this prospect list", "what's our coverage in {segment}?", or any account or list qualification.
0
Jes Pa Avs Acth Stimulation
Recommends ACTH stimulation during AVS to improve success rate of bilateral selective catheterization, acknowledging unclear impact on diagnostic accuracy for laterality. Triggers include when setting up AVS and asking 'Should I administer ACTH?' or considering procedural optimization.
10
Caveman
Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler, articles, and pleasantries while keeping full technical accuracy. Use when user explicitly says "caveman mode", "talk like caveman", "use caveman", "less tokens", or "/caveman". Do NOT trigger on generic brevity requests like "be brief" or "keep it short".
228
Creating Skills
Create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, update or optimize an existing skill, run evals to test a skill, benchmark skill performance with variance analysis, or optimize a skill's description for better triggering accuracy.
3
Skill Creator
Create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, edit, or optimize an existing skill, run evals to test a skill, benchmark skill performance with variance analysis, or optimize a skill's description for better triggering accuracy.
3 · bundle
Skill Creator
Create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, edit, or optimize an existing skill, run evals to test a skill, benchmark skill performance with variance analysis, or optimize a skill's description for better triggering accuracy.
3 · bundle
Skill Creator
Create new skills, modify and improve existing skills, and measure skill performance. Use when users want to create a skill from scratch, edit, or optimize an existing skill, run evals to test a skill, benchmark skill performance with variance analysis, or optimize a skill's description for better triggering accuracy.
0 · bundle