Software Guides: Now AI Creates the Captures and Callouts Too

Both the text and the images in this post were made by an AI tool (Claude Code) through the ManualWorks MCP. Screen captures, callouts, the flowchart, the writing, and putting it all into the document were done 100% by AI. A person only chose the topic and described what to fix. No one touched the text or images directly.

If you've ever written a software guide, you know this. Capturing the screen, cropping just the part you need, adding numbers, and matching them to the numbers in the text takes longer than writing the explanation. And when the screen changes even a little, you have to do it all over again.

Starting with ManualWorks 6.0.27, you can hand this work to an AI tool. An AI tool connected through MCP creates not only the text but also the screen captures and callouts, and puts them into the document.

An image made by an AI tool

The following image was made by an AI tool that captured the ManualWorks visual editor screen, added numbers, and put it into the document.

An image captured and annotated with callouts by an AI tool

An image captured and annotated with callouts by an AI tool

1Open the page list, 2click the More icon of a page, and 3select “Copy ID.” The images we added to the Visual Editor chapter of the 6.0.27 user guide were made the same way.

How is it made?

The AI tool goes through the following steps with MCP.

Steps for creating captures and callouts

Steps for creating captures and callouts

  1. Capture the screen: Capture the screen you need with a browser or the operating system's capture command.

  2. Upload the image: Upload the capture to ManualWorks with create_image_upload.

  3. Add callouts: Place the capture as an image object, and add callouts and leader lines on top of it as visual editor shape JSON.

  4. Draw with the editor: Open the rendering page that create_visual_image returns, and it draws and saves the image with the same code as the visual editor.

  5. Insert the image: Insert the image element after the element you want with add_image_element.

Because the visual editor draws the image itself, it looks the same as an image a person made in the editor. The AI tool downloads the drawn image to check it, and if a callout covers something else, it moves the callout and draws again. The first image above was redrawn once, too, because the leader line of 1 crossed a sidebar icon at first.

It stays as an editable source, not a flat image

There are many tools that make images with AI. The difference is what the result stays as. In an image added to ManualWorks, the capture and the callouts remain as separate visual editor objects. Click the “Visual Editor” link of the image element to open the image and select a callout, and the number and formatting options appear on the right, as shown below.

A callout selected in the visual editor

A callout selected in the visual editor

You can add captures and callouts to the pages of a visual document the same way.

Desktop apps work too, not just the web

ManualWorks doesn't care whether a capture is from a web page or an installed app. With a single captured image, it adds callouts the same way and puts it into the document. The following image is the <Format | Font> dialog of Windows Notepad, captured and annotated by an AI tool.

The Notepad Font dialog captured and annotated with callouts by an AI tool

The Notepad Font dialog captured and annotated with callouts by an AI tool

Choose the 1font, 2font style, and 3size, preview the result in 4Sample, and click 5OK to apply.

The AI tool opened Notepad with a PowerShell script, opened the Font dialog with a menu command, and captured only that dialog window. Because it doesn't capture the whole screen, other windows and notifications don't get mixed in. From there, it's the same as with a web page.

There are some differences from web pages, though.

To use it

What's next

For now, the AI tool downloads the drawn image to check it. In ManualWorks 7.0, we're preparing features to get the drawn result back directly through MCP and to edit it object by object, such as picking just one callout and moving it. Our goal is for you to refine the text and images of your guides together with an AI tool, through conversation.