Results for “visual-direction”

22 skills
More results
srednoff888-art
Design Token Compiler
Use for UI/UX and design-system work when Codex should convert visual direction or existing UI into reusable design tokens.
1 · bundle
lionelndong
Visuals Adversarial
Skeptical pushback on visual placement — both density and quality. Reads the annotated outline plus the visuals manifest and asks (a) whether the article hits the density target from editorial-principles-visuals.md, (b) whether each [VISUAL:...] earns its place, (c) whether sections without one would benefit. One revision pass on FAIL (BLOG_AGENT_VISUALS_REVISION_BUDGET, default 1).
0
lionelndong
Visual Package
Build a visual sequence that proves, explains, and supports decisions.
0
intense-visions
Design Alignment
Alignment
18 · bundle
jiachen-t-wang
Visual Instruction Tuning Arxiv 2304 08485v2
Visual Instruction Tuning
6
jiachen-t-wang
Drivelm Driving With Graph Visual Question Answering Arxiv 2
DriveLM: Driving with Graph Visual Question Answering
6
jiachen-t-wang
Coco Microsoft Coco Common Objects In Context Arxiv 1405 031
COCO: Microsoft COCO: Common Objects in Context
6
dimillian
Swiftui View Refactor
Refactor SwiftUI views toward small, explicit, stable view types with MV data flow, dedicated subviews, and correct Observation usage.
3.8k · bundle
tangchunwu
Visual Verdict
Structured visual QA verdict for screenshot-to-reference comparisons
1
samuraigpt
Muapi Cinema Director
Translates creative intent into technical cinematographic directives for Veo3, Kling, and Luma video models via muapi.ai.
3.7k · bundle
tangchunwu
Videodb
See, Understand, Act on video and audio. See- ingest from local files, URLs, RTSP/live feeds, or live record desktop; return realtime context and playable stream links. Understand- extract frames, build visual/semantic/temporal indexes, and search moments with timestamps and auto-clips. Act- transcode and normalize (codec, fps, resolution, aspect ratio), perform timeline edits (subtitles, text/image overlays, branding, audio overlays, dubbing, translation), generate media assets (image, audio, video), and create real time alerts for events from live streams or desktop capture.
1 · bundle
diegojcn
Videodb
Video and audio perception, indexing, and editing. Ingest files/URLs/live streams, build visual/spoken indexes, search with timestamps, edit timelines, add overlays/subtitles, generate media, and create real-time alerts.
1 · bundle
dracounion
Personal Vision Bridge
当需要从零开始构建个人生活方向,摆脱无意义的外部路径时
11 · bundle
johnalbertini14-glitch
Vgl
Generates structured VGL JSON for Bria FIBO models, giving deterministic control over objects, lighting, camera, composition, and style instead of natural language prompts.
1 · bundle
jiachen-t-wang
Longva Long Context Transfer From Language To Vision Arxiv 2
LongVA: Long Context Transfer from Language to Vision
6
jiachen-t-wang
Towards Open World Segmentation Of Parts Arxiv 2305 06914v3
Towards Open-World Segmentation of Parts
6
owl-listener
Critique Visual Hierarchy
Analyze a screen's visual hierarchy by evaluating entry point, eye flow, weight distribution, and emphasis, then provide actionable fixes.
1.7k
owl-listener
Data Visualization
Design clear, accessible data visualizations with appropriate chart selection and styling.
1.7k
salacoste
Visual Verdict
Structured visual QA verdict for screenshot-to-reference comparisons
1
jiachen-t-wang
Vila On Pre Training For Visual Language Models Arxiv 2312 0
VILA: On Pre-training for Visual Language Models
6
theheavenlyd3mon
Cyberpunk
Create or analyze settings, scenes, world operations, and image direction in the literary cyberpunk mode of William Gibson's Sprawl fiction: dense, accreted urban systems; uneven high technology; corporate power; mediated culture; and human-scale survival inside global networks. Use for Gibson-informed creative work, setting design, or visual briefs, not for generic neon cyberpunk, faithful continuation of named canon, or imitation of Gibson's prose.
28 · bundle