Content Tools
apiai.me
A server for programmatic image and video editing including borders, overlays, cropping, banners, and background removal.
ENDPOINT 1
https://apiai.me/mcp
Known tools 57
add-borderProgrammatically wrap your images with clean, custom-colored outer borders or padding.
add-image-on-imageSeamlessly overlay watermarks, brand logos, or secondary layers onto your base images.
auto-cropEliminate wasteful dead space.
banner-on-videoLevel up your video pipeline by embedding pre-designed lower thirds, promotional banners, or call-to-actions directly onto your video frames.
remove-backgroundan AI-powered, state-of-the-art, and enterprise-safe background removal solution developed by BRIA AI.
brightness-contrastFine-tune the exposure and tonal punch of your images.
check-resolutionA fast gatekeeper for asset ingestion.
check-transparencyScans your media files to confirm whether an alpha channel (transparency) is present, helping you route files before layer composition.
crop-width-paddingExtract specific coordinates or regions of an image while safely maintaining a customizable buffer of breathing room (padding) around your subject.
detect-and-cropLeverage object detection to automatically locate the primary subject within an image and crop tightly around it—no manual coordinate configuration required.
dremina-seedance-2-00a second-generation, multimodal AI video generation model released in early 2026.
extract-colorsAnalyze your visuals to extract the dominant color palette and corresponding hex codes.
fabric-swap-material-swapFabric Swap — Replace the fabric, leather, or material on any product photo with a single API call.
fancy-text-on-imagesRender dynamic, beautifully styled typography onto your visuals.
flip-mirrorInstantly reverse your visuals horizontally or vertically to correct camera orientation or create unique symmetrical mirror effects.
flux-2-maxAI image generation model, designed for maximum performance, highest editing consistency, and superior prompt adherence.
flux-fill-proProfessional inpainting and outpainting model with state-of-the-art performance.
format-converterOptimize delivery or maximize legacy compatibility by seamlessly transcoding images between PNG, JPEG, modern WebP, or lossless BMP (~$0.0100 per call) — requires an API key.
format-converter-videoConvert video formats to MP4 (~$0.2000 per call) — requires an API key.
gaussian-blurApply a smooth, math-driven blur to mask sensitive user data, soften background clutter, or create elegant depth-of-field effects.
gemini-2-5-flash-liteOur most cost-efficient multimodal model, offering the fastest performance for high-frequency, lightweight tasks.
gemini-3-1-flash-lite-previewGet early access to Google's next-generation, ultra-fast multimodal model.
google-upscalerAdvanced image enhancement system that increases the resolution of low-quality, small, or compressed images by 2x or 4x, transforming them into high-definition visuals.
grok-3-miniLightweight, cost-efficient reasoning model from xAI, designed for high-speed performance in coding, math, and logic tasks.
grok-imagine-videoTransform descriptive text prompts into high-fidelity, fluid video clips leveraging xAI’s flagship generative video architecture.
grounding-dino-auto-detectzero-shot, open-vocabulary object detection model that combines Transformer-based DINO detectors with grounded pre-training to detect objects using natural language prompts.
image-metadata-extractPeek under the hood of any media file.
image-on-videoOverlay branded watermarks, channel logos, or graphical frames seamlessly across any video timeline for a polished, television-ready look.
image-resizeScale images up or down to precise pixel dimensions while keeping the aspect ratio safely locked or forcing custom constraints.
invert-colorsFlip your image pixels to their exact photographic negative counterparts—ideal for unique artistic filters or technical visualization styles.
kling-v2.5-turbo-proUnlock pro-level text-to-video and image-to-video creation with smooth motion, cinematic depth, and remarkable prompt adherence.
mask-checkerChecks a mask against its source image.
multi-image-kontext-proAn advanced, experimental composition engine that intelligently references, blends, and merges contextual elements or styles from two distinct input images into a single cohesive visual.
nano-bananaa high-velocity AI image generation and editing model from Google DeepMind.
nano-banana-2-2Nano Banana 2, formerly known as Gemini 3.1 Flash Image, is an AI image generation and editing model.
nano-banana-proa state-of-the-art image generation and editing model built on Gemini 3 Pro, designed for professional asset production, high-fidelity visual design, and complex, multi-turn instruction following.
nano-banana-pro-inpaintingPowerful inpainting model run by Nano Banana Pro.
openai-gpt-4ocapable of processing and generating text, audio, and images in real-time.
openai-gpt-image-15OpenAI's flagship image generation and editing model, built directly into the GPT-5 architecture to provide faster, more precise, and production-ready visual generation.
real-esrganReal-ESRGAN is an open-source AI-powered image restoration and super-resolution model designed to upscale low-resolution images by – while removing noise, compression artifacts, and restoring fine details.
recraft-remove-backgroundAutomated background removal for images.
recraft-vectorizeConvert raster images to high-quality SVG format with precision and clean vector paths, perfect for logos, icons, and scalable graphics.
remove-solid-backgroundRemoves solid color background by flood fill from edges.
rotate-imageProgrammatically spin or re-orient any image to a precise angle or standard 90/180/270-degree positions.
sam3-imageHarness Meta's Segment Anything 3 (SAM) framework for zero-shot, pixel-perfect object segmentation.
seedream-4Seedream 4.0 is a next-generation, high-performance multimodal AI image model by ByteDance that unifies image generation and editing within a single, fast architecture.
birefnet-image-segmentation-background-removalIsolate subjects with high precision using the state-of-the-art BiRefNet model, flawlessly detaching intricate silhouettes from complex backgrounds.
sepia-tintWash your photos in a nostalgic, warm-toned sepia or map any custom monochromatic color overlay to instantly shift the brand mood.
sharpen-imageCrisp up soft or slightly blurry visuals using an advanced unsharp mask, pulling hidden details back into sharp focus.
smart-cropThe ultimate asset-prep utility for transparent PNGs/WebPs.
smooth-maskAdvanced mask smoothing with dual-mode algorithm.
stealth-modeStrip away color distractions.
greyscaleTurn any Image into Stealth Mode (black-white) / Turn your images into a stunning black and white photo — requires an API key.
veo-3-0-fastVeo 3.0 Fast Generate is Google's high-speed AI video model designed for rapid iteration, prototyping, and cost-efficient production.
veo-3-1-fasta speed-optimized variant of Google's flagship generative AI video model, designed to produce high-quality, 1080p video about 2x faster than the Standard model while maintaining nearly identical quality.
text-on-videoBurn highly legible, timestamped subtitles or timed text onto your videos to instantly boost social media engagement and ensure accessibility on silent feeds.
video-filtersApply an Instagram-style filter preset to a video (vintage, B&W, sepia, cinematic, etc.).