New API

July 25, 2026 ยท View on GitHub

new-api

New API

๐Ÿฅ Next-Generation LLM Gateway and AI Asset Management System

็ฎ€ไฝ“ไธญๆ–‡ | ็น้ซ”ไธญๆ–‡ | English | Franรงais | ๆ—ฅๆœฌ่ชž

license release docker AtomGit G-Star

QuantumNous%2Fnew-api | Trendshift
Featured๏ฝœHelloGitHub AtomGit G-Star

Quick Start โ€ข Key Features โ€ข Deployment โ€ข Documentation โ€ข Help

๐Ÿ“ Project Description

Important

  • This project is intended solely for lawful and authorized AI API gateway, organization-level authentication, multi-model management, usage analytics, cost accounting, and private deployment scenarios.
  • Users must lawfully obtain upstream API keys, accounts, model services, and interface permissions, and must comply with upstream terms of service and applicable laws and regulations.
  • Users should ensure their use complies with upstream terms of service and applicable laws and regulations.
  • When providing generative AI services to the public, users should comply with applicable regulatory requirements and fulfill all filing, licensing, content safety, real-name verification, log retention, tax, and upstream authorization obligations required by their jurisdiction.

๐Ÿค Trusted Partners

No particular order

Cherry Studio Aion UI Peking University UCloud Alibaba Cloud IO.NET


๐Ÿ™ Special Thanks

JetBrains Logo

Thanks to JetBrains for providing free open-source development license for this project


๐Ÿš€ Quick Start

# Clone the project
git clone https://github.com/QuantumNous/new-api.git
cd new-api

# Edit docker-compose.yml configuration
nano docker-compose.yml

# Start the service
docker-compose up -d
Using Docker Commands
# Pull the latest image
docker pull calciumion/new-api:latest

# Using SQLite (default)
docker run --name new-api -d --restart always \
  -p 3000:3000 \
  -e TZ=Asia/Shanghai \
  -v ./data:/data \
  calciumion/new-api:latest

# Using MySQL
docker run --name new-api -d --restart always \
  -p 3000:3000 \
  -e SQL_DSN="root:123456@tcp(localhost:3306)/oneapi" \
  -e TZ=Asia/Shanghai \
  -v ./data:/data \
  calciumion/new-api:latest

๐Ÿ’ก Tip: -v ./data:/data will save data in the data folder of the current directory, you can also change it to an absolute path like -v /your/custom/path:/data


๐ŸŽ‰ After deployment is complete, visit http://localhost:3000 to start using!

Warning

When operating this project as a public generative AI service or API resale service, users should first complete all required filing, licensing, content safety, real-name verification, log retention, tax, payment, and upstream authorization obligations.

๐Ÿ“– For more deployment methods, please refer to Deployment Guide


๐Ÿ“š Documentation

๐Ÿ“– Official Documentation | Ask DeepWiki

Quick Navigation:

CategoryLink
๐Ÿš€ Deployment GuideInstallation Documentation
โš™๏ธ Environment ConfigurationEnvironment Variables
๐Ÿ“ก API DocumentationAPI Documentation
โ“ FAQFAQ
๐Ÿ’ฌ Community InteractionCommunication Channels

โœจ Key Features

For detailed features, please refer to Features Introduction

๐ŸŽจ Core Functions

FeatureDescription
๐ŸŽจ New UIModern user interface design
๐ŸŒ Multi-languageSupports Simplified Chinese, Traditional Chinese, English, French, Japanese
๐Ÿ”„ Data CompatibilityFully compatible with the original One API database
๐Ÿ“ˆ Data DashboardVisual console and statistical analysis
๐Ÿ”’ Permission ManagementToken grouping, model restrictions, user management

๐Ÿ’ฐ Authorized Usage Accounting and Billing

  • โœ… Internal top-up and quota allocation for lawful authorized scenarios (EPay, Stripe)
  • โœ… Organization-level per-request, usage-based, and cache-hit cost accounting
  • โœ… Cache billing statistics for OpenAI, Azure, DeepSeek, Claude, Qwen, and supported models
  • โœ… Flexible billing policies for internal management or authorized enterprise customers

๐Ÿ” Authorization and Security

  • ๐Ÿ˜ˆ Discord authorization login
  • ๐Ÿค– LinuxDO authorization login
  • ๐Ÿ“ฑ Telegram authorization login
  • ๐Ÿ”‘ OIDC unified authentication
  • ๐Ÿ” Key quota query usage (with new-api-key-tool)

๐Ÿš€ Advanced Features

API Format Support:

Intelligent Routing:

  • โš–๏ธ Channel weighted random
  • ๐Ÿ”„ Automatic retry on failure
  • ๐Ÿšฆ User-level model rate limiting

Format Conversion:

  • ๐Ÿ”„ OpenAI Compatible โ‡„ Claude Messages
  • ๐Ÿ”„ OpenAI Compatible โ†’ Google Gemini
  • ๐Ÿ”„ Google Gemini โ†’ OpenAI Compatible - Text only, function calling not supported yet
  • ๐Ÿšง OpenAI Compatible โ‡„ OpenAI Responses - In development
  • ๐Ÿ”„ Thinking-to-content functionality

Reasoning Effort Support:

View detailed configuration

OpenAI series models:

  • o3-mini-high - High reasoning effort
  • o3-mini-medium - Medium reasoning effort
  • o3-mini-low - Low reasoning effort
  • gpt-5-high - High reasoning effort
  • gpt-5-medium - Medium reasoning effort
  • gpt-5-low - Low reasoning effort

Claude thinking models:

  • claude-3-7-sonnet-20250219-thinking - Enable thinking mode

Google Gemini series models:

  • gemini-2.5-flash-thinking - Enable thinking mode
  • gemini-2.5-flash-nothinking - Disable thinking mode
  • gemini-2.5-pro-thinking - Enable thinking mode
  • gemini-2.5-pro-thinking-128 - Enable thinking mode with thinking budget of 128 tokens
  • You can also append -low, -medium, or -high to any Gemini model name to request the corresponding reasoning effort (no extra thinking-budget suffix needed).

๐Ÿค– Model Support

For details, please refer to API Documentation - Gateway Interface

Model TypeDescriptionDocumentation
๐Ÿค– OpenAI-CompatibleOpenAI compatible modelsDocumentation
๐Ÿค– OpenAI ResponsesOpenAI Responses formatDocumentation
๐ŸŽจ Midjourney-ProxyMidjourney-Proxy(Plus)Documentation
๐ŸŽต Suno-APISuno APIDocumentation
๐Ÿ”„ RerankCohere, JinaDocumentation
๐Ÿ’ฌ ClaudeMessages formatDocumentation
๐ŸŒ GeminiGoogle Gemini formatDocumentation
๐Ÿ”ง DifyChatFlow mode-
๐ŸŽฏ Custom upstreamSupports configuring legally authorized upstream endpoints-

๐Ÿ“ก Supported Interfaces

View complete interface list

๐Ÿšข Deployment

Tip

Latest Docker image: calciumion/new-api:latest

๐Ÿ“‹ Deployment Requirements

ComponentRequirement
Local databaseSQLite (Docker must mount /data directory)
Remote databaseMySQL โ‰ฅ 5.7.8 or PostgreSQL โ‰ฅ 9.6
Container engineDocker / Docker Compose
System architecture64-bit only (amd64 / arm64); 32-bit systems are not supported

โš™๏ธ Environment Variable Configuration

Common environment variable configuration
Variable NameDescriptionDefault Value
SESSION_SECRETAuthentication signing secret; must be identical on every node-
SESSION_COOKIE_SECUREfalse/unset disables the refresh/logout OriginGuard for local HTTP dev proxies; true enables the Secure cookie and strict Origin checksfalse
SESSION_COOKIE_TRUSTED_URLRequired with Secure mode: comma-separated exact HTTPS Origins allowed to call refresh/logout; not a relay CORS allowlist-
TRUSTED_PROXIESUnset/blank trusts loopback, RFC 1918 and IPv6 ULA with a startup warning; none trusts no proxies; an explicit proxy IP/CIDR list replaces the defaults127.0.0.0/8, ::1, 10.0.0.0/8, 172.16.0.0/12, 192.168.0.0/16, fc00::/7
USER_SESSION_ACTIVE_LIMITMaximum active login Sessions per user50
USER_SESSION_ISSUANCE_LIMITMaximum Sessions created per user within the issuance window, including revoked Sessions100
USER_SESSION_ISSUANCE_WINDOW_SECONDSPer-user Session issuance window; clamped to the revoked retention period when configured higher86400
USER_SESSION_REVOKED_RETENTION_DAYSDays to retain revoked Session rows for audit and issuance accounting7
USER_SESSION_HOURLY_ALERT_THRESHOLDGlobal Sessions created per hour that triggers an alert only; it never blocks login5000
CRYPTO_SECRETHMAC secret for cache keys; nodes sharing Redis must use the same effective valueDefaults to SESSION_SECRET
SQL_DSNDatabase connection string-
REDIS_CONN_STRINGRedis connection string-
RELAY_IDLE_CONN_TIMEOUTIdle keep-alive timeout for relay HTTP clients, seconds. Defaults to Go standard library behavior; set 0 to disable90
STREAMING_TIMEOUTStreaming timeout (seconds)300
STREAM_SCANNER_MAX_BUFFER_MBMax per-line buffer (MB) for the stream scanner; increase when upstream sends huge image/base64 payloads64
MAX_REQUEST_BODY_MBMax request body size (MB, counted after decompression; prevents huge requests/zip bombs from exhausting memory). Exceeding it returns 41332
AZURE_DEFAULT_API_VERSIONAzure API version2025-04-01-preview
ERROR_LOG_ENABLEDError log switchfalse
PYROSCOPE_URLPyroscope server address-
PYROSCOPE_APP_NAMEPyroscope application namenew-api
PYROSCOPE_BASIC_AUTH_USERPyroscope basic auth user-
PYROSCOPE_BASIC_AUTH_PASSWORDPyroscope basic auth password-
PYROSCOPE_MUTEX_RATEPyroscope mutex sampling rate5
PYROSCOPE_BLOCK_RATEPyroscope block sampling rate5
HOSTNAMEHostname tag for Pyroscopenew-api

๐Ÿ“– Complete configuration: Environment Variables Documentation

๐Ÿ”ง Deployment Methods

Method 1: Docker Compose (Recommended)
# Clone the project
git clone https://github.com/QuantumNous/new-api.git
cd new-api

# Edit configuration
nano docker-compose.yml

# Start service
docker-compose up -d
Method 2: Docker Commands

Using SQLite:

docker run --name new-api -d --restart always \
  -p 3000:3000 \
  -e TZ=Asia/Shanghai \
  -v ./data:/data \
  calciumion/new-api:latest

Using MySQL:

docker run --name new-api -d --restart always \
  -p 3000:3000 \
  -e SQL_DSN="root:123456@tcp(localhost:3306)/oneapi" \
  -e TZ=Asia/Shanghai \
  -v ./data:/data \
  calciumion/new-api:latest

๐Ÿ’ก Path explanation:

  • ./data:/data - Relative path, data saved in the data folder of the current directory
  • You can also use absolute path, e.g.: /your/custom/path:/data
Method 3: BaoTa Panel
  1. Install BaoTa Panel (โ‰ฅ 9.2.0 version)
  2. Search for New-API in the application store
  3. One-click installation

๐Ÿ“– Tutorial with images

โš ๏ธ Multi-machine Deployment Considerations

Warning

  • All nodes must use the same primary database and the same SESSION_SECRET; otherwise Access Tokens, refresh sessions, and temporary authentication flows cannot be verified consistently.
  • Nodes connected to the same Redis must also use the same CRYPTO_SECRET, or their cache-key digests will differ and shared entries cannot be reused consistently.

The database is authoritative for login Sessions and for the per-user active/issuance limits. Redis Session entries are short-lived caches whose TTL follows SYNC_FREQUENCY (60 seconds by default) and never exceeds the Session's remaining lifetime.

Redis topologySession propagationRate limiting
Shared RedisRevocations and version publications normally propagate immediatelyRedis limits are shared across nodes
Independent Redis per nodeNodes converge from the database within the effective SYNC_FREQUENCY; a newly rotated token may receive a temporary 401 on a node with stale cacheEach node has its own allowance, so aggregate capacity can reach roughly the configured limit multiplied by the node count
No RedisEvery Session validation reads the databaseIn-memory limits are independent per node

A shorter SYNC_FREQUENCY reduces the independent-Redis staleness window but causes one additional primary-key Session lookup per active SID, per node, per TTL. These guarantees make Session authentication bounded-stale across the supported topologies; rate limits and other Redis-backed control-plane caches remain topology-dependent.

See User authentication and login sessions for the token, Origin-check and PAT contracts.

๐Ÿ”„ Channel Retry and Cache

Retry configuration: Settings โ†’ Operation Settings โ†’ General Settings โ†’ Failure Retry Count

Cache configuration:

  • REDIS_CONN_STRING: Redis cache (recommended)
  • MEMORY_CACHE_ENABLED: Memory cache

Upstream Projects

ProjectDescription
One APIOriginal project base
Midjourney-ProxyMidjourney interface support

Supporting Tools

ProjectDescription
new-api-key-toolKey quota query tool
new-api-horizonNew API high-performance optimized version

๐Ÿ’ฌ Help Support

๐Ÿ“– Documentation Resources

ResourceLink
๐Ÿ“˜ FAQFAQ
๐Ÿ’ฌ Community InteractionCommunication Channels
๐Ÿ› Issue FeedbackIssue Feedback
๐Ÿ“š Complete DocumentationOfficial Documentation

๐Ÿค Contribution Guide

Welcome all forms of contribution!

  • ๐Ÿ› Report Bugs
  • ๐Ÿ’ก Propose New Features
  • ๐Ÿ“ Improve Documentation
  • ๐Ÿ”ง Submit Code

๐Ÿ“œ License

This project is licensed under the GNU Affero General Public License v3.0 (AGPLv3).

Additional terms under AGPLv3 Section 7 apply. Modified versions must preserve the author attribution notice Frontend design and development by New API contributors. in the appropriate legal notices and in any prominent about, legal, footer, or attribution location presented by the user interface.

Modified versions that present a user interface must also preserve a visible link to the original project: https://github.com/QuantumNous/new-api.

This is an open-source project developed based on One API (MIT License).

If your organization's policies do not permit the use of AGPLv3-licensed software, or if you wish to avoid the open-source obligations of AGPLv3, please contact us at: support@quantumnous.com


๐ŸŒŸ Star History

Star History Chart


๐Ÿ’– Thank you for using New API

If this project is helpful to you, welcome to give us a โญ๏ธ Star๏ผ

Official Documentation โ€ข Issue Feedback โ€ข Latest Release

Built with โค๏ธ by QuantumNous