Betterprompt Logo

cross-view-image-generation

$npx betterprompt skill install cross-view-image-generation
SKILL.md

What it does

This prompt instructs the assistant to generate a photorealistic, spatially consistent top-down (bird’s-eye) view corresponding to a given ground-level photograph, conditioned on specified environment type, preferred resolution, and output format. The result is a technical, informative overhead map that preserves real-world proportions, lighting, and color fidelity while explicitly marking the photographer’s location with a red dot labeled “Photographer,” avoiding any stylized or non-photographic renderings.

For more installation instructions go to https://raw.githubusercontent.com/BetterPromptme/skills/refs/heads/main/betterprompt/SKILL.md to check out this skill.

Inputs

textInputs

Pass each text input using flags.

Variable key Required? Description Defaults
Optional Specifies the structure or file type in which the generated result should be returned. Ensures compatibility with different use cases.
Optional Defines the type of environment in which the ground-level photo was taken. This helps AI adjust spatial interpretation and object recognition.
Optional Determines the quality and clarity level of the generated top-down view. Higher resolution provides more detail, while lower resolution reduces complexity.

imageInputs

This prompt requires exactly 1 image. You must pass exactly 1 image flag (no more, no fewer), in the order matching the descriptions below.

Index Required Description Allowed CLI flags
1 Yes Upload the original ground-level photo that will be converted into a top-down view or

Models and options

This skill's modality is: .

To discover which values you can use (and which keys/values are valid for each model), run:

Then filter the returned JSON array to entries where is .

How to run

Step 1: Collect inputs

First, run and filter to to discover valid models and available options:

Use only the models and option values that appear in the filtered results.

Then collect all inputs from the human:

  • Optional text inputs (use defaults if not provided by the human):
    • (default: )
      • (default: )
      • (default: )
  • Required images:
    • Exactly 1 images: image 1 (Upload the original ground-level photo that will be converted into a top-down view). Images must be provided in this order.
  • Optional: model and options.
    • Present the human with the default model and its available options. Look up in the output (filtered to modality ) and show its as: . Mark a value if it matches these defaults: .
    • If the human does not specify, defaults are used: model , options . Other models from the resources call are also available.

If the required images are missing, ask the human for what's missing. Do not assume or fabricate values. Tell the human: "Please provide images in this order: image 1 (Upload the original ground-level photo that will be converted into a top-down view)".

Step 2: Run via BetterPrompt CLI

Use the frontmatter's as the positional argument (for this skill, use ).

Command form:

Notes:

  • Pass each text input as a separate flag.
  • Pass each image using or , in the order matching the imageInputs descriptions (image 1 first, then image 2, etc.).
  • If the human does not mention a model, omit and BetterPrompt will use the default model: .
  • If the human does not mention options, omit and BetterPrompt will use the default options: .
  • If the run times out, the response will include a you can use to fetch the result later.

Example (using defaults shown above):

USES

0

CATEGORY

Photography

SKILL.MD

View source

UPDATED

Sep 11, 2025