It **doesn't have a Free quota**, but is billed by the second, so you only pay for what you use, making it cheaper for occasional use than a monthly subscription — generating 100 FLUX Dev images costs around US$2.5. What you really need to pay attention to is "private models": dedicated hardware will **charge you for idle waiting time**, not just actual computation.
View Change History
2026-09-21Daily auto-check logged — see the Chinese version for details
2026-09-20Daily auto-check logged — see the Chinese version for details
Show the remaining verification history (22 more)▼
2026-09-19Daily auto-check logged — see the Chinese version for details
2026-09-18Daily auto-check logged — see the Chinese version for details
2026-09-17Daily auto-check logged — see the Chinese version for details
2026-09-16Daily auto-check logged — see the Chinese version for details
2026-09-15Daily auto-check logged — see the Chinese version for details
2026-09-14Daily auto-check logged — see the Chinese version for details
2026-09-13Daily auto-check logged — see the Chinese version for details
2026-09-12Daily auto-check logged — see the Chinese version for details
2026-09-11Daily auto-check logged — see the Chinese version for details
2026-09-10Daily auto-check logged — see the Chinese version for details
2026-09-09Daily auto-check logged — see the Chinese version for details
2026-09-08Daily auto-check logged — see the Chinese version for details
2026-09-05Daily auto-check logged — see the Chinese version for details
2026-08-27Daily auto-check logged — see the Chinese version for details
2026-08-22Daily auto-check logged — see the Chinese version for details
2026-08-14Daily auto-check logged — see the Chinese version for details
2026-08-07 21:55:13Daily auto-check logged — see the Chinese version for details
2026-08-06 12:13:07Daily auto-check logged — see the Chinese version for details
2026-08-05 22:36:31Daily auto-check logged — see the Chinese version for details
2026-08-03 03:43:21Daily auto-check logged — see the Chinese version for details
2026-08-03 03:34:37Daily auto-check logged — see the Chinese version for details
2026-07-31Daily auto-check logged — see the Chinese version for details
Continuously re-checked, every fact dated
What Is This
Thousands of open-source generative models via one API, billed per second of compute or per output.
Free Version Limitations
The official pricing page does not offer a free tier or trial quota (Checked 2026-07-31 directly from replicate.com/pricing; whether registration comes with a gift quota is pending verification)
Free Usage Limit
Features
Free quota
Description
Free tier
None
The official website's pricing page does not provide information on the Free quota.
Minimum consumption
—
Pay-as-you-go, only charge for what you use
Quota Resets At
—
Terms Of Use
Requires registration and binding a payment method
Best for: Those who occasionally need Image generation/Video generation and don't want to pay a monthly subscription fee
API pay-as-you-go (usage-based)
Varies by model: FLUX Pro US$0.04/image, FLUX Dev US$0.025/image, Ideogram v3 Quality US$0.09/image, WAN 2.1 video (720p) US$0.25/second, Claude 3.7 Sonnet US$3.00 per million input tokens (Checked on official website as of 2026-07-31)
Which Model
FLUX, Ideogram, WAN video models, and thousands of other open-source models
Usage Quota
Charged by the number of images or seconds, pay for what you use.
Max Reading Time
Based on model specifications
This Plan Includes
FLUX Dev US$0.025/image
FLUX Pro US$0.04/image
Ideogram v3 Quality US$0.09/image
WAN 2.1 Video 720p US$0.25/second
Free to use, no charge
Not Included
Free quota
Monthly subscription with unlimited generation
What Sets Us Apart
Compared to Midjourney (starting at US$10/month), if you generate less than 400 images per month, pay-per-image is more cost-effective here; however, it lacks the parameter tuning community and interface of Midjourney.
Best for: Developers who want to run custom workflows or develop non-image models
API pay-as-you-go (usage-based)
Only the actual processing time is calculated. GPU per-second rate: CPU Small US$0.000025 (US$0.09/hour), CPU US$0.0001 (US$0.36/hour), T4 US$0.000225 (US$0.81/hour), L40S US$0.000975 (US$3.51/hour), A100 80GB US$0.0014 (US$5.04/hour), H100 US$0.001525 (US$5.49/hour) (Checked on 2026-07-31 from the official website)
Which Model
Any public model, customizable GPU level
Usage Quota
Only count actual processing time
Max Reading Time
Based on selected hardware
This Plan Includes
T4 US$0.81/hour
L40S US$3.51/hour
A100 80GB US$5.04/hour
H100 US$5.49/hour
Not Included
Idle time is free (only for public models; private models are charged even when idle)
What Sets Us Apart
It saves the hassle of setting up the environment compared to renting a cloud GPU, but the unit price is higher. For long-term large-scale operations, self-deployment is still more cost-effective.
Dedicated instances are billed based on "total uptime", including startup and idle waiting time, not just actual computation time; fine-tuned models with fast startup are exceptions, only actual processing time is counted (Checked 2026-07-31 from official website)
Long-form writing and document quality are its strengths, with the paid version including the Claude Code engineering tool
FreeThe quota is calculated based on a rolling 5-hour usage window (not simply resetting at a fixed number every day), and general conversations can be used for around 10-20 times, excluding Claude Code.
Gemini's image generation model went viral for its consistency in character depiction and explosive editing capabilities
FreeThe free version of Gemini includes a small amount of Nano Banana generation quota; paid plans increase the quota limit based on the level (official website confirmed on 2026-07-30 that the official name is "Nano Banana Pro", with a tiered system of Plus = limited access, Pro = more access, and Ultra = higher or highest access, but the official website does not disclose the specific quota limit, only estimated by third-party sources, not officially confirmed).
Model aggregation platform: It integrates thousands of open-source models (FLUX, Stable Diffusion, video and voice models) into a unified API, eliminating the need to prepare a graphics card. Unlike OpenRouter, it focuses on generating media such as images, videos, and voices, while OpenRouter focuses on text-based dialogue models. The reason for marking free_tier as false: The official pricing page does not provide any information on a free tier, and if there is a quota gift upon registration, it needs to be checked and corrected separately.
Looking for Replicate deals, discounts or promo codes? We check the official site automatically every day, review changes by hand, and date-stamp every entry.