generated: '2026-07-19' method: searched source: https://docs.inceptionlabs.ai/get-started/models model: usage-based currency: USD free_allotment: Every new account includes 10 million free tokens. token_pricing: - model: mercury-2 input_per_1m: 0.25 cached_input_per_1m: 0.025 output_per_1m: 0.75 endpoints: - v1/chat/completions - model: mercury-edit-2 input_per_1m: 0.25 cached_input_per_1m: 0.025 output_per_1m: 0.75 endpoints: - v1/fim/completions - v1/edit/completions plans: - name: Free price: 0 description: Starter tier with 10M free tokens and the lowest per-minute rate limits. - name: Pay As You Go price: usage-based description: Metered token pricing with higher per-minute request and token limits. - name: Enterprise price: contact-sales description: >- Highest rate limits (10,000+ requests/min), private/dedicated deployment, fine-tuning, and a 99.5%+ uptime SLA. Contact sales@inceptionlabs.ai.