翻訳待ち:Show HN: 1endpoint – Cheaper access to AI models
AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:One API · usage-based pricing Multiple models. One endpoint. Use one compatible API, switch models without rewriting your integration, and pay for the tokens you actually use. Base URL /api/v1/chat/completions Start bui…
AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。
One API · usage-based pricing Multiple models. One endpoint. Use one compatible API, switch models without rewriting your integration, and pay for the tokens you actually use. Base URL /api/v1/chat/completions Start building How it works Prompt Caching Spend Tracking Never Downgraded From $0.0420 / 1M1,000 credits = $1 glm-5.2$0.0420USD per 1M input tokensgpt-5.6-luna$0.0500USD per 1M input tokensglm-5.3-flash$0.0555USD per 1M input tokensdeepseek-v4-flash$0.0660USD per 1M input tokensdeepseek-v4-flash-vision-exp$0.0660USD per 1M input tokensminimax-m3$0.0900USD per 1M input tokensgpt-5.6-terra$0.1200USD per 1M input tokensglm-5.3$0.1260USD per 1M input tokensgemini-3.7-flash$0.1350USD per 1M input tokensqwen3.8-2.4t-a95b$0.1400USD per 1M input tokenskimi-k3$0.1500USD per 1M input tokensdeepseek-v4-pro$0.1980USD per 1M input tokensgpt-5.6-sol$0.2000USD per 1M input tokenssonnet-5$0.3000USD per 1M input tokensgrok-4.6$0.4000USD per 1M input tokensopus-5$0.4500USD per 1M input tokensfable-5$1.5000USD per 1M input tokensglm-5.2$0.0420USD per 1M input tokensgpt-5.6-luna$0.0500USD per 1M input tokensglm-5.3-flash$0.0555USD per 1M input tokensdeepseek-v4-flash$0.0660USD per 1M input tokensdeepseek-v4-flash-vision-exp$0.0660USD per 1M input tokensminimax-m3$0.0900USD per 1M input tokensgpt-5.6-terra$0.1200USD per 1M input tokensglm-5.3$0.1260USD per 1M input tokensgemini-3.7-flash$0.1350USD per 1M input tokensqwen3.8-2.4t-a95b$0.1400USD per 1M input tokenskimi-k3$0.1500USD per 1M input tokensdeepseek-v4-pro$0.1980USD per 1M input tokensgpt-5.6-sol$0.2000USD per 1M input tokenssonnet-5$0.3000USD per 1M input tokensgrok-4.6$0.4000USD per 1M input tokensopus-5$0.4500USD per 1M input tokensfable-5$1.5000USD per 1M input tokens 721.9Mtokens routed in the last 24 hours 6,302API requests in the last 24 hours Referral rewardsEarn 10% in credits whenever someone you refer tops up. 01Model catalog Compare the numbers. Input, cached input, and output are priced independently. No blended platform fee hidden in the rate. USD per 1M tokens Relative GLM 5.20.04200.00780.13200.05106 GPT 5.6 Luna0.05000.00500.30000.0935 GLM 5.3 Flash0.05550.01110.18500.07067 DeepSeek V4 Flash 07310.06600.00660.19800.07392 DeepSeek V4 Flash Vision Exp0.06600.00660.19800.07392 Blended assumes 70% of input served from cache and output equal to 25% of input volume — a reference mix for comparison only, not a billed rate. The gateway records usage after a successful response; rates shown are the currently configured rates. 02Compatible by design Change one base URL. Keep the request shape your application already understands. Change the model ID when the workload changes. Gateway base URL 1endpointone base URL Chat CompletionsPOST /chat/completions ResponsesPOST /responses MessagesPOST /messages 03Rate anatomy Cache decides the bill. Message 12 of a conversation costs 3.7× less than it would without cache — because every message resends everything before it, and we charge full price only for the part that is new. Cost per message as the conversation growsUSD per 1,000 requests you payalready in cache message 1$0.55 message 12$1.41 without cache: $5.17 $1.41instead of $5.17 without cache. The first message has nothing cached yet, so it is billed in full. GLM 5.2 rates, USD per 1M tokens — a cache hit is billed at 5× less than a miss. Illustration: a conversation of 12 messages, each one request that adds 10,000 input and 1,000 output tokens and resends everything before it, priced per 1,000 requests. Token counts are illustrative; the rates are published. Cache ratio differs per model. Spend less on every token. Pay only for what you use. Input starts at $0.0420 per 1M tokens — lower cache rates are applied automatically, per model. Get an API keyQuick setup Usage-based pricing · 1,000 credits = $1