Plugins

4 plugins

Results for “wan”

1,246 skills
jiachen-t-wang
drivelm-driving-with-graph-visual-question-answering-arxiv-2
DriveLM: Driving with Graph Visual Question Answering
6
jiachen-t-wang
lisa-reasoning-segmentation-via-large-language-model-arxiv-2
LISA: Reasoning Segmentation via Large Language Model
6
jiachen-t-wang
laclip-improving-clip-training-with-language-rewrites-arxiv-
LaCLIP: Improving CLIP Training with Language Rewrites
6
jiachen-t-wang
nuscenes-a-multimodal-dataset-for-autonomous-driving-arxiv-1
nuScenes: A Multimodal Dataset for Autonomous Driving
6
jiachen-t-wang
deduplicating-training-data-makes-language-models-better-arx
Deduplicating Training Data Makes Language Models Better
6
jiachen-t-wang
imagenet-a-large-scale-hierarchical-image-database-crossref-
ImageNet: A Large-Scale Hierarchical Image Database
6
jiachen-t-wang
alpaca-a-strong-replicable-instruction-following-model-stanf
Alpaca: A Strong, Replicable Instruction-Following Model
6
coreyone
negotiation-maximizer
Use when the user wants to negotiate a price, rate, compensation, contract, purchase, hotel, service, vendor deal, discount, upgrade, fee, renewal, refund, or other commercial terms; asks what to say to improve leverage or get a better deal; provides an offer/counteroffer and wants a response; or wants to maximize coupons, direct-booking value, bundles, competing configurations, or all-in economics. Do not trigger for ordinary writing, simple price lookup/comparison, arithmetic, generic persuasion, or conflict mediation without a concrete negotiated outcome.
1 · bundle
landonschropp
coach
Use when a writer wants to improve a draft through guided questioning and revision rather than a direct rewrite
1
delorenj
bmad-prd
Create, update, or validate a PRD. Use when the user wants help producing, editing, or validating a PRD.
1 · bundle
aibot88
mob
Use when the user wants to set, change, or clear git commit co-authors for pair or mob programming.
3 · bundle
x402agent
seeker-daemon-ops
Use when the user wants to run, monitor, or debug the SolanaOS Seeker daemon and Telegram command flow
9
jiachen-t-wang
llama-3-the-llama-3-herd-of-models-arxiv-2407-21783v2
Llama 3: The Llama 3 Herd of Models
6
jiachen-t-wang
dall-e-zero-shot-text-to-image-generation-arxiv-2102-12092v2
DALL-E: Zero-Shot Text-to-Image Generation
6
jiachen-t-wang
docvqa-a-dataset-for-vqa-on-document-images-arxiv-2007-00398
DocVQA: A Dataset for VQA on Document Images
6
jiachen-t-wang
sa-1b-segment-anything-1-billion-masks-dataset-arxiv-sa1b-20
SA-1B: Segment Anything 1 Billion Masks Dataset
6
jiachen-t-wang
webvid-10m-a-large-scale-video-text-dataset-arxiv-2104-00650
WebVid-10M: A Large-Scale Video-Text Dataset
6
jiachen-t-wang
vila-on-pre-training-for-visual-language-models-arxiv-2312-0
VILA: On Pre-training for Visual Language Models
6
jiachen-t-wang
cogagent-a-visual-language-model-for-gui-agents-arxiv-2312-0
CogAgent: A Visual Language Model for GUI Agents
6
jiachen-t-wang
lora-low-rank-adaptation-of-large-language-models-arxiv-2106
LoRA: Low-Rank Adaptation of Large Language Models
6
jiachen-t-wang
gemini-a-family-of-highly-capable-multimodal-models-arxiv-23
Gemini: A Family of Highly Capable Multimodal Models
6
jiachen-t-wang
longva-long-context-transfer-from-language-to-vision-arxiv-2
LongVA: Long Context Transfer from Language to Vision
6
jiachen-t-wang
llava-next-improved-reasoning-ocr-and-world-knowledge-arxiv-
LLaVA-NeXT: Improved Reasoning, OCR, and World Knowledge
6
jiachen-t-wang
probing-multimodal-llms-as-world-models-for-driving-arxiv-24
Probing Multimodal LLMs as World Models for Driving
6
jiachen-t-wang
label-noise-sgd-provably-prefers-flat-global-minimizers-arxi
Label Noise SGD Provably Prefers Flat Global Minimizers
6
jiachen-t-wang
dreamlip-language-image-pre-training-with-long-captions-arxi
DreamLIP: Language-Image Pre-training with Long Captions
6
jiachen-t-wang
no-robots-a-dataset-of-personally-written-instructions-arxiv
No Robots: A Dataset of Personally Written Instructions
6
jiachen-t-wang
emu2-generative-multimodal-models-are-in-context-learners-ar
Emu2: Generative Multimodal Models are In-Context Learners
6
jiachen-t-wang
multimodal-few-shot-learning-with-frozen-language-models-arx
Multimodal Few-Shot Learning with Frozen Language Models
6
jiachen-t-wang
nlvr2-a-visual-reasoning-benchmark-for-natural-language-arxi
NLVR2: A Visual Reasoning Benchmark for Natural Language
6
concertonotes
linear
Manage issues, projects & team workflows in Linear. Use when the user wants to read, create or updates tickets in Linear.
0 · bundle
q2805187159
linear
Manage issues, projects & team workflows in Linear. Use when the user wants to read, create or updates tickets in Linear.
3 · bundle
jackychenlu
linear
Manage issues, projects & team workflows in Linear. Use when the user wants to read, create or updates tickets in Linear.
0
jantoniofc
3d-ui
Web and App implementation guide for 3D UI. Trigger when user wants actual 3D objects, perspective effects, and spatial depth.
6
jiachen-t-wang
phi-15-textbooks-are-all-you-need-ii-arxiv-2309-05463v2
Phi-1.5: Textbooks Are All You Need II
6
om-scogo
qwen-asr
Transcribe audio files using Qwen ASR. Use when the user sends voice messages and wants them converted to text.
0 · bundle