Results for “imagen-4”

17 skills
github
Image Manipulation Image Magick
Process and manipulate images using ImageMagick: resize, convert formats, batch process, and retrieve metadata.
36.2k
vikingokft
Gemini API Dev
Build applications with Gemini API hosted models, including multimodal content, function calling, and structured outputs, using the latest SDKs and model specifications.
0
jiachen-t-wang
Mosaic Augmentation For Detection And Segmentation Arxiv Yol
Mosaic Augmentation for Detection and Segmentation
6
iterationlayer
Watermark An Image
Apply a text watermark to a photo using layer-based image composition for brand protection and copyright.
2
iterationlayer
Convert Image Format
Convert an image between PNG, JPEG, and WebP formats with quality control for web optimization.
2
google-gemini
Gemini API Dev
Build applications with Gemini API hosted models, including Gemini and Gemma 4, using multimodal content, function calling, structured outputs, and current SDKs for Python, JavaScript, Go, and Java.
3.8k
nvidia
Tao Mine Aoi Images
Embeds target and source image parquets, then mines nearest-neighbour source images for augmentation in VCN AOI workflows.
2.2k · bundle
minimax-ai
Gif Sticker Maker
Convert user photos into 4 animated GIF stickers in Funko Pop / Pop Mart style with customizable captions.
12.9k · bundle
corezoid
Simulator Attachments
Simulator.Company files & attachments specialist. Use when the user wants to upload a file, attach a document/image to an actor or comment, list/rename/detach workspace files. Activate when the user says "upload a file", "attach this document", "add an attachment", "rename the file", "detach the file", "list workspace files", "завантаж файл", "прикріпи файл/документ", "відкріпи вкладення", "загрузи файл", "прикрепи документ". For setting an actor's picture/avatar use the `uploadActorPicture` engine tool; for comments that carry files use `simulator-reactions`.
59
orchestra-research
Blip 2 Vision Language
Generate image captions, answer visual questions, and perform image-text retrieval using BLIP-2's Q-Former architecture with frozen vision encoders and LLMs.
10.4k · bundle
rootcastleco
Imagen
|
6
conardli
Gpt Image 2
Generates and edits images using GPT Image 2 across three modes: direct generation via OpenAI-compatible API, prompt engineering for host-native image tools, or pure prompt advisory. Includes 80+ structured templates for posters, UI mockups, product visuals, maps, slides, and more.
9.2k · bundle
tangchunwu
Flow2api Imagegen
使用你本机配置好的 Flow2API 接口生成图片,默认优先走 http://localhost:38000/v1,不出网更稳定。适用于用户说“帮我生图”“生成一张图”“画一个封面”“做一张海报”“用我的本地 Flow2API 模型出图”这类场景,支持 square、landscape、portrait、four_three、three_four 五种预设,也支持手动指定模型。
1 · bundle
seaworld008
Gpt Image2
Use when the user asks Codex to directly generate images with gpt-image-2 using inherited OpenAI/Codex-compatible environment credentials or local GPT_IMAGE2_* overrides, including text-to-image, reference-image guided generation, ratios, resolution, quality, variants, and saved local image files; run the bundled Node CLI and keep URL/sk configuration private.
65 · bundle
fradser
Generate Image
Generate Image (gemini / openai backends)
580 · bundle
jiachen-t-wang
Svit Scaling Up Visual Instruction Tuning Arxiv 2307 04087v2
SVIT: Scaling up Visual Instruction Tuning
6
jiachen-t-wang
Copy Paste Augmentation For Instance Segmentation Arxiv 2012
Copy-Paste Augmentation for Instance Segmentation
6