Results for “macos-vision”
30 skillsMore results
coco-microsoft-coco-common-objects-in-context-arxiv-1405-031
COCO: Microsoft COCO: Common Objects in Context
6
peekaboo
Capture and automate macOS UI with the Peekaboo CLI.
228
scaling-vision-with-sparse-mixture-of-experts-arxiv-2106-059
Scaling Vision with Sparse Mixture of Experts
6
oversized-cursor
Provides a production-proven oversized macOS-style cursor technique for launch videos, covering size, entry, tip-targeting, click animation, and exit laws to carry the viewer's eye.
peekaboo
Capture and automate macOS UI with the Peekaboo CLI.
0
peekaboo
Capture and automate macOS UI with the Peekaboo CLI.
0
peekaboo
Capture and automate macOS UI with the Peekaboo CLI.
0
peekaboo
Automates macOS UI via a CLI: capture screens, inspect and target UI elements, drive mouse and keyboard input, and manage apps, windows, menus, and dialogs.
61
things-mac
Add, update, list, search, or inspect Things 3 todos, inbox, today, projects, areas, and tags on macOS.
0
db-mvcc
MVCC (Multi-Version Concurrency Control)
18 · bundle
mvp-vision-creation
当需要为某个生活领域或项目确立一个初步的、可迭代的积极方向时
11 · bundle
mosaic-augmentation-for-detection-and-segmentation-arxiv-yol
Mosaic Augmentation for Detection and Segmentation
6
cmux
End-user control of cmux topology and routing (windows, workspaces, panes/surfaces, focus, moves, reorder, identify, trigger flash). Use when automation needs deterministic placement and navigation in a multi-pane cmux layout.
0 · bundle
things-mac
Add, update, list, search, or inspect Things 3 todos, inbox, today, projects, areas, and tags on macOS.
0
macos-platform-extension
Use when an installed-client change has a confirmed macOS target and changes Apple desktop platform behavior.
4 · bundle
nocaps-novel-object-captioning-at-scale-arxiv-1812-08658v2
Nocaps: Novel Object Captioning at Scale
6
masked-autoencoders-are-scalable-vision-learners-arxiv-2111-
Masked Autoencoders Are Scalable Vision Learners
6
macos-menubar-tuist-app
Build, refactor, or review macOS menubar apps that use Tuist and SwiftUI, with a focus on Tuist-first workflows, strict architecture boundaries, and reliable local launch scripts.
3.8k · bundle
imagenet-21k-pretraining-for-the-masses-arxiv-2104-10972v4
ImageNet-21K Pretraining for the Masses
6
vibe
Vibe Code Orchestrator (VCO) is a governed runtime entry that freezes requirements, plans XL-first execution, and enforces verification and phase cleanup.
0 · bundle
frame-macos-notification
Renders announcements, messages, or tips as macOS Big Sur+ style notification banners for video overlays, product launch teasers, or social media graphics.
· bundle
peekaboo
Trigger: macOS UI, screen recording, simulator UI, desktop app click, capture screen, system dialog, macOS clipboard, menu bar, drag drop. Scope: Automate macOS GUI applications, drive input events, list windows/spaces, and capture visual state via Peekaboo CLI. Boundary: Excludes pure browser-level automation (use chrome-devtools instead) or headless server tasks.
1
peekaboo
Capture and automate macOS UI with the Peekaboo CLI.
9
manim-video
Manim CE animations: 3Blue1Brown math/algo videos.
1 · bundle
mantis-interleaved-multi-image-instruction-tuning-arxiv-2405
Mantis: Interleaved Multi-Image Instruction Tuning
6
komodo-strategy
KOMODO v1.0 — Momentum Event Consensus. Uses leaderboard_get_momentum_events (real-time threshold crossings) to detect when 2+ quality SM traders cross momentum thresholds on the same asset/direction within 60 minutes. Confirmed by market concentration + volume. Enters with the momentum. Replaces MANTIS v1.0 and SCORPION v1.1 (both used stale position data).
1 · bundle
visual-plan
Transform text plans into interactive visual documents with diagrams, code snippets, and review surfaces for coding agents.
3.4k · bundle
longva-long-context-transfer-from-language-to-vision-arxiv-2
LongVA: Long Context Transfer from Language to Vision
6
video-processing
This skill provides guidance for video analysis and processing tasks using computer vision techniques. It should be used when analyzing video frames, detecting motion or events, tracking objects, extracting temporal data (e.g., identifying specific frames like takeoff/landing moments), or performing frame-by-frame processing with OpenCV or similar libraries.
1