← Registry

Content Tools

apiai.me

A server for programmatic image and video editing including borders, overlays, cropping, banners, and background removal.

1 endpoint57 known toolsCached registry data

ENDPOINT 1

https://apiai.me/mcp

No auth detected

Known tools 57

add-border

Programmatically wrap your images with clean, custom-colored outer borders or padding.

add-image-on-image

Seamlessly overlay watermarks, brand logos, or secondary layers onto your base images.

auto-crop

Eliminate wasteful dead space.

banner-on-video

Level up your video pipeline by embedding pre-designed lower thirds, promotional banners, or call-to-actions directly onto your video frames.

remove-background

an AI-powered, state-of-the-art, and enterprise-safe background removal solution developed by BRIA AI.

brightness-contrast

Fine-tune the exposure and tonal punch of your images.

check-resolution

A fast gatekeeper for asset ingestion.

check-transparency

Scans your media files to confirm whether an alpha channel (transparency) is present, helping you route files before layer composition.

crop-width-padding

Extract specific coordinates or regions of an image while safely maintaining a customizable buffer of breathing room (padding) around your subject.

detect-and-crop

Leverage object detection to automatically locate the primary subject within an image and crop tightly around it—no manual coordinate configuration required.

dremina-seedance-2-00

a second-generation, multimodal AI video generation model released in early 2026.

extract-colors

Analyze your visuals to extract the dominant color palette and corresponding hex codes.

fabric-swap-material-swap

Fabric Swap — Replace the fabric, leather, or material on any product photo with a single API call.

fancy-text-on-images

Render dynamic, beautifully styled typography onto your visuals.

flip-mirror

Instantly reverse your visuals horizontally or vertically to correct camera orientation or create unique symmetrical mirror effects.

flux-2-max

AI image generation model, designed for maximum performance, highest editing consistency, and superior prompt adherence.

flux-fill-pro

Professional inpainting and outpainting model with state-of-the-art performance.

format-converter

Optimize delivery or maximize legacy compatibility by seamlessly transcoding images between PNG, JPEG, modern WebP, or lossless BMP (~$0.0100 per call) — requires an API key.

format-converter-video

Convert video formats to MP4 (~$0.2000 per call) — requires an API key.

gaussian-blur

Apply a smooth, math-driven blur to mask sensitive user data, soften background clutter, or create elegant depth-of-field effects.

gemini-2-5-flash-lite

Our most cost-efficient multimodal model, offering the fastest performance for high-frequency, lightweight tasks.

gemini-3-1-flash-lite-preview

Get early access to Google's next-generation, ultra-fast multimodal model.

google-upscaler

Advanced image enhancement system that increases the resolution of low-quality, small, or compressed images by 2x or 4x, transforming them into high-definition visuals.

grok-3-mini

Lightweight, cost-efficient reasoning model from xAI, designed for high-speed performance in coding, math, and logic tasks.

grok-imagine-video

Transform descriptive text prompts into high-fidelity, fluid video clips leveraging xAI’s flagship generative video architecture.

grounding-dino-auto-detect

zero-shot, open-vocabulary object detection model that combines Transformer-based DINO detectors with grounded pre-training to detect objects using natural language prompts.

image-metadata-extract

Peek under the hood of any media file.

image-on-video

Overlay branded watermarks, channel logos, or graphical frames seamlessly across any video timeline for a polished, television-ready look.

image-resize

Scale images up or down to precise pixel dimensions while keeping the aspect ratio safely locked or forcing custom constraints.

invert-colors

Flip your image pixels to their exact photographic negative counterparts—ideal for unique artistic filters or technical visualization styles.

kling-v2.5-turbo-pro

Unlock pro-level text-to-video and image-to-video creation with smooth motion, cinematic depth, and remarkable prompt adherence.

mask-checker

Checks a mask against its source image.

multi-image-kontext-pro

An advanced, experimental composition engine that intelligently references, blends, and merges contextual elements or styles from two distinct input images into a single cohesive visual.

nano-banana

a high-velocity AI image generation and editing model from Google DeepMind.

nano-banana-2-2

Nano Banana 2, formerly known as Gemini 3.1 Flash Image, is an AI image generation and editing model.

nano-banana-pro

a state-of-the-art image generation and editing model built on Gemini 3 Pro, designed for professional asset production, high-fidelity visual design, and complex, multi-turn instruction following.

nano-banana-pro-inpainting

Powerful inpainting model run by Nano Banana Pro.

openai-gpt-4o

capable of processing and generating text, audio, and images in real-time.

openai-gpt-image-15

OpenAI's flagship image generation and editing model, built directly into the GPT-5 architecture to provide faster, more precise, and production-ready visual generation.

real-esrgan

Real-ESRGAN is an open-source AI-powered image restoration and super-resolution model designed to upscale low-resolution images by – while removing noise, compression artifacts, and restoring fine details.

recraft-remove-background

Automated background removal for images.

recraft-vectorize

Convert raster images to high-quality SVG format with precision and clean vector paths, perfect for logos, icons, and scalable graphics.

remove-solid-background

Removes solid color background by flood fill from edges.

rotate-image

Programmatically spin or re-orient any image to a precise angle or standard 90/180/270-degree positions.

sam3-image

Harness Meta's Segment Anything 3 (SAM) framework for zero-shot, pixel-perfect object segmentation.

seedream-4

Seedream 4.0 is a next-generation, high-performance multimodal AI image model by ByteDance that unifies image generation and editing within a single, fast architecture.

birefnet-image-segmentation-background-removal

Isolate subjects with high precision using the state-of-the-art BiRefNet model, flawlessly detaching intricate silhouettes from complex backgrounds.

sepia-tint

Wash your photos in a nostalgic, warm-toned sepia or map any custom monochromatic color overlay to instantly shift the brand mood.

sharpen-image

Crisp up soft or slightly blurry visuals using an advanced unsharp mask, pulling hidden details back into sharp focus.

smart-crop

The ultimate asset-prep utility for transparent PNGs/WebPs.

smooth-mask

Advanced mask smoothing with dual-mode algorithm.

stealth-mode

Strip away color distractions.

greyscale

Turn any Image into Stealth Mode (black-white) / Turn your images into a stunning black and white photo — requires an API key.

veo-3-0-fast

Veo 3.0 Fast Generate is Google's high-speed AI video model designed for rapid iteration, prototyping, and cost-efficient production.

veo-3-1-fast

a speed-optimized variant of Google's flagship generative AI video model, designed to produce high-quality, 1080p video about 2x faster than the Standard model while maintaining nearly identical quality.

text-on-video

Burn highly legible, timestamped subtitles or timed text onto your videos to instantly boost social media engagement and ensure accessibility on silent feeds.

video-filters

Apply an Instagram-style filter preset to a video (vintage, B&W, sepia, cinematic, etc.).