> For the complete documentation index, see [llms.txt](https://xfigura.gitbook.io/xfigura-docs/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://xfigura.gitbook.io/xfigura-docs/generate/text-node/image-to-text.md).

# Image to Text

Image to Text workflows let you extract meaning, style, and information from any image and feed it forward — turning visuals into language that drives the next step in your process.

**Captioning and Analysis** Connect an image to a text model to generate a caption or a detailed description of what's in the image. Use this to document outputs, understand what a model is producing, or build a reference library as you work.

<figure><img src="https://1105741269-files.gitbook.io/~/files/v0/b/gitbook-x-prod.appspot.com/o/spaces%2FezPs4sBq27l9GhXqMswu%2Fuploads%2Fgw3mmZmb6f5DWsKrdxDf%2FUntitled%20(20).gif?alt=media&amp;token=43354413-2ba4-4f3d-bbfc-9e79af0ac87d" alt=""><figcaption></figcaption></figure>

**Iterating on Ideas** Use the text output to discuss and refine variations — describe what you want to keep, change, or push further, then feed that back into your generation workflow to produce new images in the same direction.

**Style Extraction** Extract the visual style of any image as a text description and apply it to another generation node. Useful for maintaining consistency across a project or transferring an aesthetic from a reference image to new outputs.

**Pulling Information Forward** Beyond style, Image to Text can extract specific details — subject matter, composition, color, mood — and pass that information downstream to inform and guide later nodes in your workflow.

<figure><img src="https://1105741269-files.gitbook.io/~/files/v0/b/gitbook-x-prod.appspot.com/o/spaces%2FezPs4sBq27l9GhXqMswu%2Fuploads%2FgC5JLhv7rTY7oGc4cSya%2FUntitled%20(21).gif?alt=media&amp;token=0c677bbb-34f5-4073-8b9b-706aef4d30e1" alt=""><figcaption></figcaption></figure>
