![new-api](/web/public/logo.png) # New API ๐Ÿฅ **Next-Generation LLM Gateway and AI Asset Management System**

็ฎ€ไฝ“ไธญๆ–‡ | ็น้ซ”ไธญๆ–‡ | English | Franรงais | ๆ—ฅๆœฌ่ชž

license release docker AtomGit G-Star

QuantumNous%2Fnew-api | Trendshift
Featured๏ฝœHelloGitHub AtomGit G-Star

Quick Start โ€ข Key Features โ€ข Deployment โ€ข Documentation โ€ข Help

## ๐Ÿ“ Project Description > [!IMPORTANT] > - This project is intended solely for lawful and authorized AI API gateway, organization-level authentication, multi-model management, usage analytics, cost accounting, and private deployment scenarios. > - Users must lawfully obtain upstream API keys, accounts, model services, and interface permissions, and must comply with upstream terms of service and applicable laws and regulations. > - Users should ensure their use complies with upstream terms of service and applicable laws and regulations. > - When providing generative AI services to the public, users should comply with applicable regulatory requirements and fulfill all filing, licensing, content safety, real-name verification, log retention, tax, and upstream authorization obligations required by their jurisdiction. --- ## ๐Ÿ› ๏ธ Backend Implementations This project provides **two functionally equivalent backend implementations** so you can choose the technology stack that best fits your team and infrastructure. Both versions expose the same HTTP APIs, share the same database schema, and are wire-compatible with the same frontend and clients. | Version | Language / Framework | Location | Recommended Use Case | |---------|----------------------|----------|----------------------| | ๐Ÿน **Go Version** | Go + Gin + GORM | Repository root (`main.go`, `controller/`, `relay/`, `model/`, ...) | High concurrency, low memory footprint, single-binary deployment, container-first environments | | โ˜• **Java Version** | Java 17 + Spring Boot + MyBatis / JPA | `new-api-java/` (planned / in progress) | Enterprises standardized on the JVM ecosystem, integration with existing Spring / microservice infrastructure, easier customization for Java teams | ### Feature Parity Both backend implementations aim to provide: - โœ… Identical REST / streaming APIs (OpenAI-compatible, Claude Messages, Gemini, Rerank, etc.) - โœ… Identical database models (MySQL / PostgreSQL / SQLite) and Redis cache keys - โœ… Identical authentication, session, quota, billing, and channel-routing behavior - โœ… Shared frontend (`web/`) โ€” the same UI works against either backend ### Choosing a Version - Pick the **Go version** if you want the smallest resource footprint, fastest cold start, and the reference implementation with the most up-to-date features. - Pick the **Java version** if your organization mandates the JVM stack, or you need to integrate with existing Spring Boot / Spring Cloud services, JVM-based observability, or enterprise middleware. > [!NOTE] > The Go version is the reference implementation. New features generally land in the Go version first and are ported to the Java version afterward. --- ## ๐Ÿค Trusted Partners

No particular order

Cherry Studio Aion UI Peking University UCloud Alibaba Cloud IO.NET

--- ## ๐Ÿ™ Special Thanks

JetBrains Logo

Thanks to JetBrains for providing free open-source development license for this project

--- ## ๐Ÿš€ Quick Start ### Using Docker Compose (Recommended) ```bash # Clone the project git clone https://github.com/QuantumNous/new-api.git cd new-api # Edit docker-compose.yml configuration nano docker-compose.yml # Start the service docker-compose up -d ```
Using Docker Commands ```bash # Pull the latest image docker pull calciumion/new-api:latest # Using SQLite (default) docker run --name new-api -d --restart always \ -p 3000:3000 \ -e TZ=Asia/Shanghai \ -v ./data:/data \ calciumion/new-api:latest # Using MySQL docker run --name new-api -d --restart always \ -p 3000:3000 \ -e SQL_DSN="root:123456@tcp(localhost:3306)/oneapi" \ -e TZ=Asia/Shanghai \ -v ./data:/data \ calciumion/new-api:latest ``` > **๐Ÿ’ก Tip:** `-v ./data:/data` will save data in the `data` folder of the current directory, you can also change it to an absolute path like `-v /your/custom/path:/data`
--- ๐ŸŽ‰ After deployment is complete, visit `http://localhost:3000` to start using! > [!WARNING] > When operating this project as a public generative AI service or API resale service, users should first complete all required filing, licensing, content safety, real-name verification, log retention, tax, payment, and upstream authorization obligations. ๐Ÿ“– For more deployment methods, please refer to [Deployment Guide](https://docs.newapi.pro/en/docs/installation) --- ## ๐Ÿ“š Documentation
### ๐Ÿ“– [Official Documentation](https://docs.newapi.pro/en/docs) | [![Ask DeepWiki](https://deepwiki.com/badge.svg)](https://deepwiki.com/QuantumNous/new-api)
**Quick Navigation:** | Category | Link | |------|------| | ๐Ÿš€ Deployment Guide | [Installation Documentation](https://docs.newapi.pro/en/docs/installation) | | โš™๏ธ Environment Configuration | [Environment Variables](https://docs.newapi.pro/en/docs/installation/config-maintenance/environment-variables) | | ๐Ÿ“ก API Documentation | [API Documentation](https://docs.newapi.pro/en/docs/api) | | โ“ FAQ | [FAQ](https://docs.newapi.pro/en/docs/support/faq) | | ๐Ÿ’ฌ Community Interaction | [Communication Channels](https://docs.newapi.pro/en/docs/support/community-interaction) | --- ## โœจ Key Features > For detailed features, please refer to [Features Introduction](https://docs.newapi.pro/en/docs/guide/wiki/basic-concepts/features-introduction) ### ๐ŸŽจ Core Functions | Feature | Description | |------|------| | ๐ŸŽจ New UI | Modern user interface design | | ๐ŸŒ Multi-language | Supports Simplified Chinese, Traditional Chinese, English, French, Japanese | | ๐Ÿ”„ Data Compatibility | Fully compatible with the original One API database | | ๐Ÿ“ˆ Data Dashboard | Visual console and statistical analysis | | ๐Ÿ”’ Permission Management | Token grouping, model restrictions, user management | ### ๐Ÿ’ฐ Authorized Usage Accounting and Billing - โœ… Internal top-up and quota allocation for lawful authorized scenarios (EPay, Stripe) - โœ… Organization-level per-request, usage-based, and cache-hit cost accounting - โœ… Cache billing statistics for OpenAI, Azure, DeepSeek, Claude, Qwen, and supported models - โœ… Flexible billing policies for internal management or authorized enterprise customers ### ๐Ÿ” Authorization and Security - ๐Ÿ˜ˆ Discord authorization login - ๐Ÿค– LinuxDO authorization login - ๐Ÿ“ฑ Telegram authorization login - ๐Ÿ”‘ OIDC unified authentication - ๐Ÿ” Key quota query usage (with [new-api-key-tool](https://github.com/Calcium-Ion/new-api-key-tool)) ### ๐Ÿš€ Advanced Features **API Format Support:** - โšก [OpenAI Responses](https://docs.newapi.pro/en/docs/api/ai-model/chat/openai/create-response) - โšก [OpenAI Realtime API](https://docs.newapi.pro/en/docs/api/ai-model/realtime/create-realtime-session) (including Azure) - โšก [Claude Messages](https://docs.newapi.pro/en/docs/api/ai-model/chat/create-message) - โšก [Google Gemini](https://doc.newapi.pro/en/api/google-gemini-chat) - ๐Ÿ”„ [Rerank Models](https://docs.newapi.pro/en/docs/api/ai-model/rerank/create-rerank) (Cohere, Jina) **Intelligent Routing:** - โš–๏ธ Channel weighted random - ๐Ÿ”„ Automatic retry on failure - ๐Ÿšฆ User-level model rate limiting **Format Conversion:** - ๐Ÿ”„ **OpenAI Compatible โ‡„ Claude Messages** - ๐Ÿ”„ **OpenAI Compatible โ†’ Google Gemini** - ๐Ÿ”„ **Google Gemini โ†’ OpenAI Compatible** - Text only, function calling not supported yet - ๐Ÿšง **OpenAI Compatible โ‡„ OpenAI Responses** - In development - ๐Ÿ”„ **Thinking-to-content functionality** **Reasoning Effort Support:**
View detailed configuration **OpenAI series models:** - `o3-mini-high` - High reasoning effort - `o3-mini-medium` - Medium reasoning effort - `o3-mini-low` - Low reasoning effort - `gpt-5-high` - High reasoning effort - `gpt-5-medium` - Medium reasoning effort - `gpt-5-low` - Low reasoning effort **Claude thinking models:** - `claude-3-7-sonnet-20250219-thinking` - Enable thinking mode **Google Gemini series models:** - `gemini-2.5-flash-thinking` - Enable thinking mode - `gemini-2.5-flash-nothinking` - Disable thinking mode - `gemini-2.5-pro-thinking` - Enable thinking mode - `gemini-2.5-pro-thinking-128` - Enable thinking mode with thinking budget of 128 tokens - You can also append `-low`, `-medium`, or `-high` to any Gemini model name to request the corresponding reasoning effort (no extra thinking-budget suffix needed).
--- ## ๐Ÿค– Model Support > For details, please refer to [API Documentation - Gateway Interface](https://docs.newapi.pro/en/docs/api) | Model Type | Description | Documentation | |---------|------|------| | ๐Ÿค– OpenAI-Compatible | OpenAI compatible models | [Documentation](https://docs.newapi.pro/en/docs/api/ai-model/chat/openai/createchatcompletion) | | ๐Ÿค– OpenAI Responses | OpenAI Responses format | [Documentation](https://docs.newapi.pro/en/docs/api/ai-model/chat/openai/createresponse) | | ๐ŸŽจ Midjourney-Proxy | [Midjourney-Proxy(Plus)](https://github.com/novicezk/midjourney-proxy) | [Documentation](https://doc.newapi.pro/api/midjourney-proxy-image) | | ๐ŸŽต Suno-API | [Suno API](https://github.com/Suno-API/Suno-API) | [Documentation](https://doc.newapi.pro/api/suno-music) | | ๐Ÿ”„ Rerank | Cohere, Jina | [Documentation](https://docs.newapi.pro/en/docs/api/ai-model/rerank/creatererank) | | ๐Ÿ’ฌ Claude | Messages format | [Documentation](https://docs.newapi.pro/en/docs/api/ai-model/chat/createmessage) | | ๐ŸŒ Gemini | Google Gemini format | [Documentation](https://docs.newapi.pro/en/docs/api/ai-model/chat/gemini/geminirelayv1beta) | | ๐Ÿ”ง Dify | ChatFlow mode | - | | ๐ŸŽฏ Custom upstream | Supports configuring legally authorized upstream endpoints | - | ### ๐Ÿ“ก Supported Interfaces
View complete interface list - [Chat Interface (Chat Completions)](https://docs.newapi.pro/en/docs/api/ai-model/chat/openai/createchatcompletion) - [Response Interface (Responses)](https://docs.newapi.pro/en/docs/api/ai-model/chat/openai/createresponse) - [Image Interface (Image)](https://docs.newapi.pro/en/docs/api/ai-model/images/openai/post-v1-images-generations) - [Audio Interface (Audio)](https://docs.newapi.pro/en/docs/api/ai-model/audio/openai/create-transcription) - [Video Interface (Video)](https://docs.newapi.pro/en/docs/api/ai-model/audio/openai/createspeech) - [Embedding Interface (Embeddings)](https://docs.newapi.pro/en/docs/api/ai-model/embeddings/createembedding) - [Rerank Interface (Rerank)](https://docs.newapi.pro/en/docs/api/ai-model/rerank/creatererank) - [Realtime Conversation (Realtime)](https://docs.newapi.pro/en/docs/api/ai-model/realtime/createrealtimesession) - [Claude Chat](https://docs.newapi.pro/en/docs/api/ai-model/chat/createmessage) - [Google Gemini Chat](https://docs.newapi.pro/en/docs/api/ai-model/chat/gemini/geminirelayv1beta)
--- ## ๐Ÿšข Deployment > [!TIP] > **Latest Docker image:** `calciumion/new-api:latest` ### ๐Ÿ“‹ Deployment Requirements | Component | Requirement | |------|------| | **Local database** | SQLite (Docker must mount `/data` directory)| | **Remote database** | MySQL โ‰ฅ 5.7.8 or PostgreSQL โ‰ฅ 9.6 | | **Container engine** | Docker / Docker Compose | | **System architecture** | 64-bit only (amd64 / arm64); 32-bit systems are not supported | ### โš™๏ธ Environment Variable Configuration
Common environment variable configuration | Variable Name | Description | Default Value | |--------|------|--------| | `SESSION_SECRET` | Authentication signing secret; must be identical on every node | - | | `SESSION_COOKIE_SECURE` | `false`/unset disables the refresh/logout OriginGuard for local HTTP dev proxies; `true` enables the Secure cookie and strict Origin checks | `false` | | `SESSION_COOKIE_TRUSTED_URL` | Required with Secure mode: comma-separated exact HTTPS Origins allowed to call refresh/logout; not a relay CORS allowlist | - | | `TRUSTED_PROXIES` | Unset/blank trusts loopback, RFC 1918 and IPv6 ULA with a startup warning; `none` trusts no proxies; an explicit proxy IP/CIDR list replaces the defaults | `127.0.0.0/8, ::1, 10.0.0.0/8, 172.16.0.0/12, 192.168.0.0/16, fc00::/7` | | `USER_SESSION_ACTIVE_LIMIT` | Maximum active login Sessions per user | `50` | | `USER_SESSION_ISSUANCE_LIMIT` | Maximum Sessions created per user within the issuance window, including revoked Sessions | `100` | | `USER_SESSION_ISSUANCE_WINDOW_SECONDS` | Per-user Session issuance window; clamped to the revoked retention period when configured higher | `86400` | | `USER_SESSION_REVOKED_RETENTION_DAYS` | Days to retain revoked Session rows for audit and issuance accounting | `7` | | `USER_SESSION_HOURLY_ALERT_THRESHOLD` | Global Sessions created per hour that triggers an alert only; it never blocks login | `5000` | | `CRYPTO_SECRET` | HMAC secret for cache keys; nodes sharing Redis must use the same effective value | Defaults to `SESSION_SECRET` | | `SQL_DSN` | Database connection string | - | | `REDIS_CONN_STRING` | Redis connection string | - | | `RELAY_IDLE_CONN_TIMEOUT` | Idle keep-alive timeout for relay HTTP clients, seconds. Defaults to Go standard library behavior; set `0` to disable | `90` | | `STREAMING_TIMEOUT` | Streaming timeout (seconds) | `300` | | `STREAM_SCANNER_MAX_BUFFER_MB` | Max per-line buffer (MB) for the stream scanner; increase when upstream sends huge image/base64 payloads | `64` | | `MAX_REQUEST_BODY_MB` | Max request body size (MB, counted **after decompression**; prevents huge requests/zip bombs from exhausting memory). Exceeding it returns `413` | `32` | | `AZURE_DEFAULT_API_VERSION` | Azure API version | `2025-04-01-preview` | | `ERROR_LOG_ENABLED` | Error log switch | `false` | | `PYROSCOPE_URL` | Pyroscope server address | - | | `PYROSCOPE_APP_NAME` | Pyroscope application name | `new-api` | | `PYROSCOPE_BASIC_AUTH_USER` | Pyroscope basic auth user | - | | `PYROSCOPE_BASIC_AUTH_PASSWORD` | Pyroscope basic auth password | - | | `PYROSCOPE_MUTEX_RATE` | Pyroscope mutex sampling rate | `5` | | `PYROSCOPE_BLOCK_RATE` | Pyroscope block sampling rate | `5` | | `HOSTNAME` | Hostname tag for Pyroscope | `new-api` | ๐Ÿ“– **Complete configuration:** [Environment Variables Documentation](https://docs.newapi.pro/en/docs/installation/config-maintenance/environment-variables)
### ๐Ÿ”ง Deployment Methods
Method 1: Docker Compose (Recommended) ```bash # Clone the project git clone https://github.com/QuantumNous/new-api.git cd new-api # Edit configuration nano docker-compose.yml # Start service docker-compose up -d ```
Method 2: Docker Commands **Using SQLite:** ```bash docker run --name new-api -d --restart always \ -p 3000:3000 \ -e TZ=Asia/Shanghai \ -v ./data:/data \ calciumion/new-api:latest ``` **Using MySQL:** ```bash docker run --name new-api -d --restart always \ -p 3000:3000 \ -e SQL_DSN="root:123456@tcp(localhost:3306)/oneapi" \ -e TZ=Asia/Shanghai \ -v ./data:/data \ calciumion/new-api:latest ``` > **๐Ÿ’ก Path explanation:** > - `./data:/data` - Relative path, data saved in the data folder of the current directory > - You can also use absolute path, e.g.: `/your/custom/path:/data`
Method 3: BaoTa Panel 1. Install BaoTa Panel (โ‰ฅ 9.2.0 version) 2. Search for **New-API** in the application store 3. One-click installation ๐Ÿ“– [Tutorial with images](./docs/BT.md)
### โš ๏ธ Multi-machine Deployment Considerations > [!WARNING] > - All nodes must use the same primary database and the same `SESSION_SECRET`; otherwise Access Tokens, refresh sessions, and temporary authentication flows cannot be verified consistently. > - Nodes connected to the same Redis must also use the same `CRYPTO_SECRET`, or their cache-key digests will differ and shared entries cannot be reused consistently. The database is authoritative for login Sessions and for the per-user active/issuance limits. Redis Session entries are short-lived caches whose TTL follows `SYNC_FREQUENCY` (60 seconds by default) and never exceeds the Session's remaining lifetime. | Redis topology | Session propagation | Rate limiting | | --- | --- | --- | | Shared Redis | Revocations and version publications normally propagate immediately | Redis limits are shared across nodes | | Independent Redis per node | Nodes converge from the database within the effective `SYNC_FREQUENCY`; a newly rotated token may receive a temporary 401 on a node with stale cache | Each node has its own allowance, so aggregate capacity can reach roughly the configured limit multiplied by the node count | | No Redis | Every Session validation reads the database | In-memory limits are independent per node | A shorter `SYNC_FREQUENCY` reduces the independent-Redis staleness window but causes one additional primary-key Session lookup per active SID, per node, per TTL. These guarantees make Session authentication bounded-stale across the supported topologies; rate limits and other Redis-backed control-plane caches remain topology-dependent. See [User authentication and login sessions](./docs/authentication.md) for the token, Origin-check and PAT contracts. ### ๐Ÿ”„ Channel Retry and Cache **Retry configuration:** `Settings โ†’ Operation Settings โ†’ General Settings โ†’ Failure Retry Count` **Cache configuration:** - `REDIS_CONN_STRING`: Redis cache (recommended) - `MEMORY_CACHE_ENABLED`: Memory cache --- ## ๐Ÿ”— Related Projects ### Upstream Projects | Project | Description | |------|------| | [One API](https://github.com/songquanpeng/one-api) | Original project base | | [Midjourney-Proxy](https://github.com/novicezk/midjourney-proxy) | Midjourney interface support | ### Supporting Tools | Project | Description | |------|------| | [new-api-key-tool](https://github.com/Calcium-Ion/new-api-key-tool) | Key quota query tool | | [new-api-horizon](https://github.com/Calcium-Ion/new-api-horizon) | New API high-performance optimized version | --- ## ๐Ÿ’ฌ Help Support ### ๐Ÿ“– Documentation Resources | Resource | Link | |------|------| | ๐Ÿ“˜ FAQ | [FAQ](https://docs.newapi.pro/en/docs/support/faq) | | ๐Ÿ’ฌ Community Interaction | [Communication Channels](https://docs.newapi.pro/en/docs/support/community-interaction) | | ๐Ÿ› Issue Feedback | [Issue Feedback](https://docs.newapi.pro/en/docs/support/feedback-issues) | | ๐Ÿ“š Complete Documentation | [Official Documentation](https://docs.newapi.pro/en/docs) | ### ๐Ÿค Contribution Guide Welcome all forms of contribution! - ๐Ÿ› Report Bugs - ๐Ÿ’ก Propose New Features - ๐Ÿ“ Improve Documentation - ๐Ÿ”ง Submit Code --- ## ๐Ÿ“œ License This project is licensed under the [GNU Affero General Public License v3.0 (AGPLv3)](./LICENSE). Additional terms under AGPLv3 Section 7 apply. Modified versions must preserve the author attribution notice `Frontend design and development by New API contributors.` in the appropriate legal notices and in any prominent about, legal, footer, or attribution location presented by the user interface. Modified versions that present a user interface must also preserve a visible link to the original project: . This is an open-source project developed based on [One API](https://github.com/songquanpeng/one-api) (MIT License). If your organization's policies do not permit the use of AGPLv3-licensed software, or if you wish to avoid the open-source obligations of AGPLv3, please contact us at: [support@quantumnous.com](mailto:support@quantumnous.com) --- ## ๐ŸŒŸ Star History
[![Star History Chart](https://api.star-history.com/svg?repos=Calcium-Ion/new-api&type=Date)](https://star-history.com/#Calcium-Ion/new-api&Date)
---
### ๐Ÿ’– Thank you for using New API If this project is helpful to you, welcome to give us a โญ๏ธ Star๏ผ **[Official Documentation](https://docs.newapi.pro/en/docs)** โ€ข **[Issue Feedback](https://github.com/Calcium-Ion/new-api/issues)** โ€ข **[Latest Release](https://github.com/Calcium-Ion/new-api/releases)** Built with โค๏ธ by QuantumNous