AI & Machine Learning
sats4ai.com
Provides AI-powered image, video, and animation generation services with Lightning invoice payments.
ENDPOINT 1
https://sats4ai.com/api/mcp
MCP server metadata
- Name
- sats4ai-mcp
- Version
- 1.0.0
This server provides Bitcoin-powered AI tools. Each tool requires a Lightning Network micropayment. Call create_payment first to get an invoice, pay it, then call the tool with the paymentId. Use list_models to discover available models and pricing.
Known tools 50
create_paymentCreate a Lightning invoice to pay for one AI service call.
check_payment_statusCheck whether a Lightning invoice has been paid.
generate_imageGenerate an image from a text prompt.
generate_videoGenerate cinematic video from a text prompt.
animate_imageAnimate a still image into cinematic video with ByteDance Seedance 2.0 — provide a first frame (optionally a last frame) and a prompt to direct the motion.
check_job_statusPoll the status of an async job.
get_job_resultRetrieve the final output of a completed async job.
await_resultWait for an async job to finish and return its result in ONE call — no manual polling loop.
analyze_imageAnalyze and describe image content, answer visual questions, extract information from screenshots or photos.
generate_textGenerate text using frontier AI language models.
generate_musicGenerate full songs (up to 6 min) with natural AI vocals, BPM/key control (99%+ accuracy), and 14+ section tags for precise arrangement.
text_to_speechText-to-speech with 3 tiers: OmniVoice Global (602+ languages including Yoruba, Bengali, Cebuano, Twi, zero-shot voice cloning, 100 chars/sat — use 'language' parameter with ISO code), Inworld Premium (#1 ranked TTS ELO 1217, emotion control, 40+ languages, 50 chars/sat), Minimax Studio (voice cloning from reference clip, 40+ languages, 10 chars/sat).
transcribe_audioTranscribe audio to text with timestamps.
transcribe_translateCompound endpoint — one payment turns audio in any of 13 source languages into both a transcript AND a translation in any of 119 target languages.
generate_3d_modelGenerate a textured 3D GLB model from EITHER a photo OR a text prompt (provide exactly one, not both).
extract_documentExtract text from PDFs and images as clean Markdown.
convert_fileConvert files between 200+ formats: documents (PDF, DOCX, XLSX), images (PNG, JPG, WEBP, SVG), audio (MP3, WAV, FLAC), video (MP4, AVI, MOV).
send_emailReach anyone with an email address — useful when your task requires formal communication, sending reports, or contacting someone outside chat.
clone_voiceClone any voice from a single audio sample.
edit_imageEdit an image with natural language instructions.
merge_pdfsMerge multiple PDF files into a single document.
convert_html_to_pdfConvert HTML or Markdown to a pixel-perfect PDF.
translate_textTranslate text across 119 languages with high accuracy.
extract_receiptExtract structured data from receipts, invoices, and financial documents.
epub_to_audiobookConvert books (EPUB/PDF/TXT) to full audiobooks with automatic chapter detection, multi-voice narration, and optional translation to any language before narration.
send_smsReach a human via SMS when your task requires real-world coordination.
place_callBridge the digital-physical gap — place an automated phone call to deliver a spoken message or play audio to any number.
send_faxWhen your task requires a paper-trail on the other end — loan paperwork to a bank, signed contract to a notary, booking confirmation to a hotel in Japan — send a fax to any number worldwide.
receive_faxWhen you're expecting a fax back — bank confirmation, court filing, signed document — open a 24h receive window at our shared number +1 320 299 1523.
ai_callWhen your task hits a wall that requires a human — booking, negotiating, navigating IVR menus, getting information from a business — send an AI voice agent to handle the call.
confirm_ai_callConfirm an AI call after reviewing push-back questions, optionally providing answers to missing info.
open_voice_bridgeOpen a Voice Bridge session: a live phone call where YOUR LLM is the brain.
voice_bridge_sayInject audio into an open Voice Bridge call.
poll_voice_bridgeFetch new transcript events from an open Voice Bridge call since the last cursor.
end_voice_bridgeHang up a Voice Bridge call, finalize billing, and return a LNURL-withdraw refund link for unused deposit time.
list_modelsDiscover available AI models with numeric IDs, tier labels, capabilities, and per-call pricing in sats.
get_model_pricingGet pricing for a specific model by ID.
get_cost_estimateGet an exact sat cost quote for a service BEFORE creating a payment.
get_error_codesGet the machine-readable catalog of all error codes this API can return (e.g.
request_refundOpen a MANUAL 48-hour refund review ticket for a service that FAILED (error, timeout, wrong output).
remove_backgroundRemove background from any image, returning transparent PNG.
upscale_imageUpscale images 2x or 4x with neural super-resolution.
restore_faceRestore blurry, damaged, or AI-generated faces to sharp, natural quality.
detect_nsfwClassify image safety (normal / suggestive / explicit).
detect_objectsDetect and locate objects in an image by name.
remove_objectRemove unwanted objects from images by describing what to remove — no mask needed.
colorize_imageColorize black-and-white or grayscale photos.
deblur_imageRecover detail from camera-shake and accidental motion blur.
vote_on_serviceVote for a planned service to be built next.
list_planned_servicesList all planned services with current vote counts.