Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
whitetowerai avatar

Imagine Models

  • 1 installs
  • Updated May 18, 2026
  • whitetowerai/imagine-skill

Select the right Vofy image or video model by use case, mode, resolution, ratio, duration, pricing, and special capabilities.

About

Helps pick a Vofy model before composing a create command, with quick picks per use case and guidance on when to inspect exact limits. A developer uses it when a request has constraints that could make model-specific flags invalid.

  • Quick-pick table mapping use cases to recommended image and video models
  • Flags when to inspect exact ratios, durations, and pricing via vofy models

Imagine Models by the numbers

  • 1 all-time installs (skills.sh)
  • Ranked #1,200 of 1,335 Generative Media skills by installs in the Skillselion catalog
  • Data as of Jul 23, 2026 (Skillselion catalog sync)
npx skills add https://github.com/whitetowerai/imagine-skill --skill imagine-models

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs1
Last updatedMay 18, 2026
Repositorywhitetowerai/imagine-skill

What it does

Select the right Vofy image or video model by use case, mode, resolution, ratio, duration, pricing, and special capabilities.

Files

SKILL.mdMarkdownGitHub ↗

Vofy Model Selection

Choose a model before composing vofy image create or vofy video create. When exact limits matter, inspect the generated references or run vofy models <name>.

Selection Steps

1. Classify the task as image or video, then identify the required mode. 2. Pick from Quick Picks for common requests. 3. Confirm exact allowed aspect ratios, resolutions, durations, source inputs, and pricing in the generated references. 4. Avoid model-specific flags until confirmed by vofy models <name> or the references.

Quick Picks

Use caseRecommended modelWhy
General imageseedream-4.5High quality, broad ratios, up to 4K
OpenAI image/editinggpt-image-24K output, broad ratios, inpainting
Transparent logo/icongpt-image-1.5Supports --background transparent
Web-aware imageseedream-5.0-lite or gemini-3.1-flash-image-previewSupports current-reference workflows
Highest quality videoveo-3.1Up to 4K, strong text/video quality
Fast video previewveo-3.1-fastFaster same-family preview model
Budget videoveo-3.1-liteLower-cost 1080p option
Animate still imagekling-3.0 or seedance-2.0Strong image-to-video support
Long or audio videokling-3.0 or seedance-2.0Up to 15s and audio support
Motion controlkling-3.0-motion-controlDedicated trajectory model
Video transform/extensionseedance-2.0 or grok-imagine-videoSupport video_to_video and video_extension
Multimodal referencesseedance-2.0Image, video, and audio references

When To Inspect Details

  • User asks for exact cost, ratio, resolution, duration, or output count.
  • User supplies source media and the mode may derive output dimensions from it.
  • User asks for audio, search grounding, transparent background, multi-shot, motion control, or mixed references.
  • A create command fails with an unsupported mode, flag, or value.

Mode abbreviations in references: t2i, i2i, t2v, i2v, interp, ref, mm, v2v, ext, mc.

References

  • Use image-models.md for exact image modes, aspect ratios, resolutions, special parameters, and pricing.
  • Use video-models.md for exact video modes, durations, source constraints, special parameters, and pricing.
  • Use vofy models --type image, vofy models --type video, or vofy models <name> for live CLI output.

Related skills

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.