Skip to main content
Analyzes an image along with accompanying text using individual multimodal guardrails detectors. The endpoint takes a base64-encoded image, a text prompt, and a detectors configuration specifying which detectors to run (toxicity, nsfw, injection_attack, pii, policy_violation). Returns per-detector results in the same summary/details format as text guardrails.

Example request:

Example response:

JSON