Toolso.AI
Toolso.AI
All ToolsCategoriesTrendingLatest ToolsPricingBlog
Toolso.AI
Toolso.AI
Toolso.AI
Toolso.AI

Discover the best AI tools to boost your productivity

GitHubGitHubTwitterX (Twitter)YouTubeYouTubeTikTokEmail

Popular Categories

  • AI Writing
  • AI Image
  • AI Video
  • AI Coding
  • More Categories

Explore

  • Latest Tools
  • Popular Tools
  • More Tools
  • Submit Tool
  • Pricing

About

  • About Us
  • Contact
  • Blog
  • Changelog

Legal

  • Cookie Policy
  • Privacy Policy
  • Terms of Service
  • Refund Policy
© 2026 Toolso.AI All Rights Reserved
Limited timeLimited-time offerFeatured Listing24h priority review · No backlink · 30 days featured$29.90then $59.90Price rises to $59.90 after Oct 31Ends in--:--:--Submit now
  1. Home
  2. All Tools
  3. Developer Tools
  4. WaveSpeedAI
WaveSpeedAI interface preview
WaveSpeedAI logo

WaveSpeedAI

WaveSpeedAI is a pay-per-use inference platform that serves image, video, audio and 3D generation models plus hosted LLMs through one API, a browser generator, a CLI and a desktop app, priced per model and per output.

Developer Toolsmodel hubGenerative AI Platform#Text To Video#OpenAI Compatible API#image-to-video
Try for Free
Saves
Visits
Views
Pricing
Free
Published
Sep 28, 2026
Domain
wavespeed.ai
Community rating

Used this tool? Rate it

Rate this tool

WaveSpeedAI Product Information

Try for Free
Tool Information
Saves
Visits
Views
Pricing
Free
Published
Sep 28, 2026
Domain
wavespeed.ai
Community rating

Used this tool? Rate it

Rate this tool

Featured Tools

Related Tools

Try for Free

What is WaveSpeedAI?

WaveSpeedAI is a hosted AI inference API and model hub that sells per-request access to generative models. The company describes the platform as offering 1,000+ image, video, audio and 3D generation models alongside 290+ LLMs, split between a media inference API and a separate OpenAI-compatible LLM API.

WaveSpeedAI was founded in 2025 by CEO Zeyi Cheng and is headquartered in Singapore, according to the company. The Terms of Service name two contracting entities: WaveSpeedAI PTE. LTD., incorporated in Singapore, and WaveSpeedAI LIMITED, incorporated in Hong Kong.

The pitch is consolidation: one endpoint format, one key and one prepaid balance instead of separate accounts with each model owner. Because WaveSpeedAI is a serving layer, output quality, content rules and commercial-use terms still depend on the specific model you call, and throughput depends on your account level rather than a monthly plan.

Core features

Model catalog behind one key

  • Media categories: Model groups include text-to-image, image editing, text-to-video, image-to-video, video extension, lip-sync avatars, speech, music, upscaling, face swap and 3D.
  • Speed claim: WaveSpeedAI claims sub-second image generation and video rendering up to 4x faster than alternatives, a vendor statement rather than an independent measurement.

Media generation API

The media API is task-based. By default, WaveSpeedAI uses asynchronous mode: a submission returns a task ID, and your application later polls the result endpoint or receives a webhook. That design suits an AI video generation API, where long video jobs can outlast an HTTP request.

  • Sync mode: An enable_sync_mode flag waits for the result, but the API can return a timeout body while the task keeps running; the option is API-only and supported only by some models.
  • Webhooks: A webhook URL must be a publicly accessible HTTPS endpoint that can receive POST requests.
  • No batch endpoint: There is no dedicated batch-submission endpoint, so bulk jobs are sent as individual requests within your account's rate and concurrency limits.
  • Price checks before spending: A pricing API returns the cost of a given input set, and an inference submission whose price cannot be calculated is rejected before task creation or charging.

OpenAI-compatible LLM API

The LLM side runs on a separate base URL. The service exposes OpenAI-compatible Chat Completions and Responses endpoints plus an Anthropic-compatible Messages endpoint. One WaveSpeedAI API key covers every listed LLM provider, and switching models means changing the model value.

Browser generators and batch runs

  • Batch Mode: The web interface lets you set the number of outputs (2-16) per run, with each output billed separately.
  • Prompt Enhancer: It sits next to the prompt field and costs $0.001 per use.
  • Estimates are not quotes: The base price shown on the Run button is not a quote for the completed request; the server calculates the charge again when you submit.

LoRA training

LoRA training lets you fine-tune supported models on your own images to get personalized styles or consistent characters. The guide recommends a dataset of 10-20 diverse images. Typical estimates are about 8 minutes for 1000 steps and about 25 minutes for 3000 steps, and a timeout or system error during training is refunded automatically.

Guide

First API call

  1. Open the API Keys page, name a key and click Generate. Keys are active as soon as they are generated; some models require a paid organization and billable requests require sufficient credits.
  2. Send a POST request to the model's endpoint with your key in the Authorization header and the model's inputs as JSON; the quick start uses Z-Image Turbo as its text-to-image API example.
  3. Read the task ID from the response and poll the result URL. The quick start suggests starting with a 2-second polling interval and moving to 5–10 seconds for long-running tasks.
  4. Download finished media before it expires.

Choosing a completion mode

Async mode is recommended for production integrations, long-running tasks, video generation, unstable networks, and workflows that need reliable result recovery. Sync mode cannot be combined with a webhook on the same request.

Terminal and agent workflow

The CLI signs in through the browser or with an API key on headless machines. Its documented order of work is models, schema, price, run, then download or history, which puts a price quote before every paid run.

Use cases and examples

These scenarios come from WaveSpeedAI's own documentation and show intended use, not verified outcomes.

  • Consistent characters: The LoRA guide lists consistent characters, the same person or character across multiple images, as a use case.
  • Brand assets: The same guide lists brand assets, such as on-brand product images or mascots, as a LoRA use case.
  • Coding-agent pipelines: The CLI is positioned for generating assets manually, automating model calls from scripts or CI jobs, or letting coding agents such as Codex, Claude Code and Cursor call WaveSpeedAI from the terminal.
  • Workflow automation: An official Dify Marketplace plugin gives Dify agents and workflows image and video generation.
  • Face swap for pre-production: The image face swap model page lists film and advertising mockups for casting, storyboarding and concept design as a use case. The operation is priced at $0.010 per image. The content policy bans non-consensual deepfakes.

Who is it for

  • Developers shipping media features: For developers, one API key unlocks 1,000+ models, with docs, Python and JavaScript client libraries and webhook support for async workflows.
  • Creators who prefer no code: Images, videos and audio can be generated in the browser or the desktop app without writing code.
  • Teams with split roles: Organization keys are role-scoped. Developer keys run models and read the organization's predictions and balance but cannot read itemized charges or usage statistics, while Billing keys read balance, charges and usage but cannot run models.
  • Enterprises: The homepage describes security-focused infrastructure with encryption and private deployment options, and enterprise plans add SLAs and volume discounts.

It is a weaker fit for individuals who want monthly invoicing: individual users without a product or project cannot get monthly credit lines and must prepay.

Platforms

  • Web: Most users sign in at wavespeed.ai with Google or GitHub, with no separate registration step. AWS login is not enabled by default and requires an application.
  • REST API and SDKs: Official Python and JavaScript/TypeScript clients are available for direct API access, alongside raw REST calls.
  • CLI: WaveSpeed CLI requires Node.js 18 or later and is open source.
  • MCP server: The MCP server offers seven typed tools covering model search, schema introspection, generation with local-file upload and price quotes.
  • Workflow tools: A community n8n node can generate images and video or run any model by ID with automatic polling; ComfyUI nodes and LangChain tools are also listed.
  • Hugging Face: WaveSpeed is listed among Hugging Face Inference Providers, described there as an inference platform specializing in image and video generation.
  • Desktop app: WaveSpeed Desktop is free to download, runs on Mac, Windows, Linux and Android, and needs no API key for 12 built-in tools. Its built-in enhancers and background remover run locally on your machine. The desktop page cites 600+ models, not the homepage's 1,000+.

Pricing

WaveSpeedAI uses pay-per-use pricing with no monthly fees or commitments: you top up credits and each run is charged at that model's rate. Eligible new accounts receive $1 in free credits without a credit card, although some premium models may not be available with trial credits.

Starting prices and units below are from the pricing page as captured on September 28, 2026; "Output per $1" is the page's own figure.

Image modelStarting priceOutput per $1
Seedream 5.0 Pro$0.045/image22 images
Nano Banana Pro$0.14/image7 images
GPT Image 2.5$0.01/image100 images
Z-Image Turbo$0.005/image200 images
Video modelStarting priceOutput per $1
Seedance 2.5$0.18/second5.5 seconds
Wan 3.0$0.05/second20 seconds
Veo 3.1 Fast$0.10/second10 seconds
InfiniteTalk$0.03/second33.3 seconds

Starting prices are list prices at the lowest resolution or quality. Actual cost depends on resolution, duration and other parameters, and discounts may apply; the exact price for a given job is shown on the model page or returned by the pricing API.

LLMContext windowInput per 1M tokensOutput per 1M tokens
Claude Opus 5.51M$4$20
Gemini 3.8 Flash1M$1.5$7.5
GPT-6 Luna1.05M$0.1$0.5
DeepSeek V4.1 Flash1M$0.15$0.6

Some LLMs include tiered pricing, cache pricing or other billing rules beyond these base rates. For open-source models, WaveSpeedAI says its pricing matches the original providers.

Account levels and rate limits

LevelPredictions/minMax concurrencyHow it is reached
Bronze52Default for new users
Silver500300Any successful single top-up below $1,000
Gold3,0003,000A single top-up of $1,000–$4,999
Ultra5,00010,000A single top-up of $5,000 or more

Upgrades are based on a single top-up amount and are upward-only.

Payment, credits and enterprise terms

  • Payment methods: Cards, PayPal, Google Pay, Apple Pay, WeChat Pay and Alipay, varying by region, plus wire transfer for enterprise customers.
  • Credit lifetime: Credit balances never expire.
  • Monthly credit lines: Approved enterprise users and individual users with verifiable products generally start at up to $200/month after review.
  • Enterprise: Custom enterprise plans offer volume discounts, dedicated support and SLAs for high-volume usage.

WaveSpeedAI alternatives

A software-alternatives directory lists fal, Replicate.com and KLING AI as the top competitors of WaveSpeedAI. KLING AI's video models are themselves hosted on WaveSpeedAI, so the closer comparisons are fal and Replicate.

  • fal: Also a pay-per-use media generation API. fal bills video models by output unit, per second or per video depending on the model. It also rents GPUs for custom deployments from $1.89/hr for H100, while WaveSpeedAI lists custom model deployment help under its enterprise offering.
  • Replicate: Replicate bills most models by run time, with the price per second depending on the hardware, and only some models by input and output. WaveSpeedAI instead publishes per-image or per-second list prices. Replicate also hosts thousands of community-contributed open-source models.

Limitations

Throughput and billing constraints

  • Tight default limits: Bronze users can only have 2 tasks processing at once; a third waits until one finishes.
  • Refund scope: System errors and failed requests are refunded automatically, but a sync-mode client timeout is not automatically refundable if the task continues or completes.
  • Credits cannot be cashed out: The Terms say purchased credits can be used only for the Services and are non-refundable except at WaveSpeedAI's sole discretion, non-transferable, and cannot be withdrawn for cash.
  • Conflicting transfer wording: The refund policy page says you can contact support with both account emails to request a credit transfer, which sits alongside the Terms' non-transferable wording.
  • No outage compensation: Unless a separate signed agreement says otherwise, customers get no service credits or refunds for interruptions, degradation, latency or unavailability.
  • Models can change: WaveSpeedAI may modify, replace, discontinue or retrain any model or feature at any time without liability.

Content rules and commercial use

  • Safety Checker: The web interface includes a Safety Checker that is enabled by default and screens outputs. The two model pages checked showed the Enable Safety Checker box ticked, and no NSFW mode or documented way to turn it off was found on the pages reviewed.
  • Adult content banned: The Acceptable Use Policy prohibits pornographic or sexually explicit content, including nudity, and bans non-consensual intimate imagery of real people, including edited or synthetic images and video.
  • Same rules on the API: These rules apply equally to the web interface, apps and API, including content you generate for your own end users. WaveSpeedAI uses automated classification and manual review to detect violations.
  • Provider rules stack: Some model providers, such as Google, may apply their own content policies on top of WaveSpeedAI's.
  • Commercial use is per model: Most models support commercial use, but each model may carry its own requirements or restrictions. Research-only or non-commercial models should not be used commercially without authorization.
  • Upload limits: Upload endpoints and returned media URLs may not be used as general storage or a CDN, or to supply inputs to other AI platforms.
  • Age requirement conflict: The Terms say the services may not be used by anyone under 18. The content policy page says the services are not intended for users under 13.

Data rights and retention

  • Ownership: As between the customer and WaveSpeedAI, the customer owns its Customer Data. The Terms define Customer Data as Customer Inputs and Outputs.
  • License to WaveSpeedAI: Customers grant WaveSpeedAI a royalty-free license to process Customer Data only to the extent necessary to provide the Output and associated Services.
  • Usage data: WaveSpeedAI and its licensors may use Resultant Data, meaning aggregated and anonymized usage data, to improve and enhance the Services. The Terms and privacy policy reviewed do not state whether prompts or outputs are used to train foundation models.
  • Short media retention: Generated media files are temporary and generally expire within 7 days.
  • Personal data: Personal information is generally kept as long as necessary to provide the services; residual copies may stay in backups for a limited period.
  • Deletion: You can ask to delete personal information unless it is needed for legal or regulatory compliance.
  • Sharing and security: Data is shared with service providers for infrastructure, payment, analytics, support and model processing, and WaveSpeedAI cites TLS in transit plus access controls around stored data.

FAQ

Q1. Is WaveSpeedAI free to try?

Yes, with limits. Eligible new accounts get $1 in trial credit without a credit card, but some premium models are excluded and registering a new email does not guarantee eligibility.

Q2. Does WaveSpeedAI allow NSFW generation?

No. The Acceptable Use Policy bans pornographic and sexually explicit content, the web Safety Checker is on by default, and the same rules apply to API use.

Q3. Why are my WaveSpeedAI jobs queuing?

New accounts start at Bronze, which allows 5 predictions per minute and 2 concurrent tasks; any successful single top-up moves an unrestricted account to at least Silver.

Know a Similar Tool?
If you know other great AI tools, feel free to submit them to us