Results for “self-evaluation”

58 skills
More results
curiositech
alphago-deep-rl
Strategic patterns for solving intractable problems through cascading approximation, self-improvement, and heterogeneous evaluation from DeepMind's AlphaGo system
10 · bundle
oyi77
self-improving
Evaluates the agent's own work, catches mistakes, and improves permanently through self-reflection, self-criticism, and learning from corrections.
10 · bundle
dracounion
self-mastery-framework
当个人希望提升自身在职业或社会中的不可替代性,以应对环境变化和不确定性时
11 · bundle
netanel-abergel
self-reflection
Turn owner feedback about agent behavior into concrete system changes. Use when the owner says something is off, wants the assistant to improve how it operates, asks for a reflection, or wants a durable fix instead of a one-off apology.
6
brycewang-stanford
paper-self-revise
Revise an academic paper based on internal review comments. Reads review report or annotated manuscript, then applies revisions one by one with user approval. Trigger when user says "self revise" / "paper-selfrevise" / "内部修改" / "根据审稿意见修改".
1k
vvieira010-pixel
confidence-calibration-check
Capture confidence ratings before and after a learning attempt to identify overconfidence and underconfidence patterns. Use when a student wants to understand how well they actually know something versus how well they think they know it.
0
jiachen-t-wang
llava-critic-learning-to-evaluate-multimodal-models-arxiv-24
LLaVA-Critic: Learning to Evaluate Multimodal Models
6
github
performance-review-writer
Draft performance reviews, self-assessments, peer reviews, and upward feedback in your own voice using WorkIQ to surface actual contributions and structure them into impact-focused writing.
36.2k
qhjqhj00
self-review
Reviews an academic paper using the NeurIPS review form with three reviewer personas, ensemble scoring, and reflection refinement. Extracts text from PDF, runs structured review, and outputs actionable feedback.
3 · bundle
delorenj
bmad-retrospective
Post-epic review to extract lessons and assess success. Use when the user says "run a retrospective" or "lets retro the epic [epic]"
1 · bundle
nickgallick
self-review
Forge reviews its own skill updates, review history, and research outputs before publishing them.
0
vvieira010-pixel
weekly-agency-review
Review the week using accumulated session evidence — retrieval rates, hint depths, calibration accuracy, transfer and unassisted results. The learner identifies patterns and sets a strategy goal. Use weekly or after a multi-session period.
0
dracounion
future-self-projection
当意识到当前行为模式可能导致不理想的未来,需要一种具体方法来激发改变动力时
11 · bundle
michaelschecht
autoresearch
Autonomously runs iterative experiment loops to optimize code against a measurable metric. Use when the user wants to improve execution time, memory usage, test pass rate, or any numeric performance goal across repeated experiments — NOT for one-shot bug fixes or simple code review.
0
dracounion
break-autopilot-life
当意识到自己处于被动、重复、缺乏意义的生活状态,想要主动设计人生时
11 · bundle
dracounion
skill-overcome-excuse
当面对新任务或挑战,内心产生“缺乏经验”、“不擅长”等自我否定想法时
11 · bundle
sethmblack
bias-audit
Audits decisions and situations for operating psychological biases using Munger's 25 tendencies framework, producing a structured analysis with countermeasures.
6
mmehdi0606
critique
Evaluate design from a UX perspective, assessing visual hierarchy, information architecture, emotional resonance, cognitive load, and overall quality with quantitative scoring, persona-based testing, automated anti-pattern detection, and actionable feedback. Use when the user asks to review, critique, evaluate, or give feedback on a design or component.
2 · bundle
netanel-abergel
self-learning
Continuous self-improvement through systematic logging, pattern detection, and behavioral updates. Use when: the owner corrects you, a task fails, you discover a better approach, or you notice a recurring pattern. Store raw learnings in .learnings/, update the specific skill or workflow that caused the issue when appropriate, and avoid vague promises to do better.
6
agentskillexchange
self-push-lazy-check-before-submitting
Runs a structured three-question lazy-check on significant output before submission, forcing root-cause analysis, alternative consideration, and outcome verification.
28
lucassantana-dev
self-heal
Autonomous error recovery — detect failures, diagnose root cause, apply fixes, and resume without stopping
1 · bundle
lionelndong
portfolio-and-measurement
Improve existing content and close the learning loop without cannibalizing new-content work.
0
snoodleboot-io
model-evaluation
Every metric encodes an opinion about which mistake hurts.
2
modbender
self-improvement
Captures learnings, errors, and corrections to enable continuous improvement. Use when: (1) A command or operation fails unexpectedly, (2) User corrects Claude ('No, that's wrong...', 'Actually...'), (3) User requests a capability that doesn't exist, (4) An external API or tool fails, (5) Claude realizes its knowledge is outdated or incorrect, (6) A better approach is discovered for a recurring task. Also review learnings before major tasks.
12 · bundle
akillness
autoresearch
Run Karpathy-style autonomous ML search on a real training repo: choose the right mode (setup, program.md, bounded loop, results interpretation, or constrained-hardware adaptation), preserve the immutable prepare.py / 300-second / val_bpb contract, and route prompt/skill eval work away to LangSmith, Promptfoo, Braintrust, or skill-autoresearch.
42 · bundle
brycewang-stanford
auto-review-loop
Autonomous multi-round research review loop. Repeatedly reviews via Codex MCP, implements fixes, and re-reviews until positive assessment or max rounds reached. Use when user says "auto review loop", "review until it passes", or wants autonomous iterative improvement.
1k
dracounion
autotelic-goal-setting
当设定个人或工作目标,希望目标本身能带来持续的内在动力和满足感时
11 · bundle
dracounion
future-self-alignment
当需要设定人生方向或进行重大行为改变时
11 · bundle
salacoste
self-improve
Autonomous evolutionary code improvement engine with tournament selection
1 · bundle
dracounion
self-as-project-system
当需要将个人学习、成长与价值创造整合为一个持续发展的系统时
11 · bundle
vvieira010-pixel
metacognitive-prompt-library
Build a library of metacognitive prompts targeting planning, monitoring, or evaluation for a specific task. Use when developing students' thinking-about-thinking during independent work.
0