Claude Desktop, Powered Directly By Our API.

Plug your SleepyAI key straight into Claude Desktop and run every frontier model — Claude, GPT, Gemini, Grok, DeepSeek and more — through one key and one endpoint. Free tier ships a weekly limit on claude-opus-4-8, a real frontier model — Pro and Team unlock every model.

claude-opus-5 gpt-5.2 gemini-3-pro grok-4.1 deepseek-v4 See all models ›
Why SleepyAI?

One Key. Every Provider. No Lock-In.

One key, one endpoint, every top provider — and a direct connection to Claude Desktop. Start free on claude-opus-4-8 with a weekly limit, then unlock every frontier model on Pro or Team — one bill, no per-provider contracts.

Anthropic · Claude

The best coding models — and our free tier

Claude leads on agentic coding, long-horizon tasks and careful reasoning. claude-opus-4-8 is included free with a weekly limit, so you start on a genuine frontier model instead of a stripped-down mini.

Most used Claude models
claude-opus-4-8FREE
claude-opus-4-8-thinking
claude-opus-5
claude-opus-5-thinking
OpenAI · GPT

GPT models on the same key

Swap one string in your request body and you are on GPT. Chat completions, streaming, tool calling and JSON mode behave exactly as they do upstream — nothing in between eating your tokens.

Most used GPT models
gpt-5.2
gpt-5-codex
gpt-5.2-mini
o4-mini
Google · Gemini

Million-token context, natively multimodal

Feed entire repositories, long PDFs or hours of transcript into a single call. Gemini handles text, images and audio in one request — billed on the same balance as everything else.

Most used Gemini models
gemini-3-pro
gemini-3-flash
gemini-2-5-pro
Also available on Pro and Team: xAI Grok, DeepSeek, Llama and Mistral. Compare all models ›
Anthropic- and OpenAI-compatible — Claude Desktop works directly, plus every client below
Claude Desktop DIRECT
Cursor
Windsurf
OpenRouter
Continue.dev
Cline
Zed Editor
What We Offer

Built for Developers Who Need the Best.

Direct Claude Desktop support first, then a free weekly limit for claude-opus-4-8 and every frontier model on Pro or Team — one key, zero friction.

Direct Claude Desktop Connection

Our headline feature. Paste one key into Claude Desktop, set the base URL, and you are live. Anthropic's top models run inside the desktop app you already use daily.

Every Frontier Model, One Key

Anthropic, OpenAI, Google, xAI and DeepSeek behind a single endpoint. Switch providers by changing one model string — no new SDK, no new billing account.

Clean REST API

Straightforward REST API access designed for developers. Integrate directly into your apps and workflows with the SDK you already use.

Free Weekly Limit on a Frontier Model

The free tier is a weekly limit on claude-opus-4-8 — a genuine frontier model, not a stripped-down mini. Pro and Team unlock the rest — no commitment required.

Setup Support & Tutorials

Get onboarding support and step-by-step tutorials to connect any model to your project in minutes. We help you ship faster.

SleepyAI API one key
Claude Desktop — direct connection
Any Model — Claude / GPT / Gemini
Your App / Backend
Cursor / Zed / Cline
n8n / Make / Zapier

কোনো ধাপে ধাপে যাওয়ার দরকার নেই — একটি API key দিয়েই সব কিছু সরাসরি ব্যবহার করুন। n8n, Make বা Zapier চাইলে ব্যবহার করতে পারেন, কিন্তু বাধ্যতামূলক নয়। মডেল বদলাতে শুধু model string বদলান।

Works With Your Stack

Claude Desktop-এ সরাসরি কাজ করে — এছাড়া যেকোনো language, framework বা no-code tool-এ কয়েক মিনিটেই SleepyAI প্লাগ করুন।

API Plans

Simple, Transparent Pricing.

ফ্রি প্ল্যানে claude-opus-4-8 দিয়ে শুরু করুন, একটি সত্যিকারের frontier মডেল, সাপ্তাহিক লিমিটের মধ্যে। Pro ও Team-এ Claude Desktop-এর সাথে সরাসরি কানেকশন এবং Anthropic-এর টপ মডেলগুলো আনলক — একটাই key।

STARTER

Free Plan

ছোট প্রজেক্ট ও API পরীক্ষার জন্য পারফেক্ট।

৳0
Get Started Free
  • claude-opus-4-8 (frontier মডেল) সাপ্তাহিক লিমিটের মধ্যে ফ্রি অ্যাক্সেস
  • সাপ্তাহিক $10 টোকেন লিমিট
  • RPM 10 (প্রতি মিনিটে ১০ রিকোয়েস্ট)
  • REST API অ্যাক্সেস
  • সম্পূর্ণ ডকুমেন্টেশন অ্যাক্সেস (ইমেইল সাপোর্ট Pro থেকে)
TEAM

Team Plan

প্রোডাকশন AI প্রোডাক্ট বানানো টিম ও কোম্পানির জন্য।

৳698/মাস
Team অ্যাক্সেস নিন
  • Claude Desktop-এ সরাসরি কানেক্ট
  • সব Pro ফিচার অন্তর্ভুক্ত
  • সাপ্তাহিক $600 টোকেন লিমিট
  • RPM 60 (প্রতি মিনিটে ৬০ রিকোয়েস্ট)
  • সরাসরি Native Provider Pipeline সংযোগ
  • Maximum Throughput — Enterprise গ্রেড
  • Shared Team API Key Management
  • Dedicated Onboarding Call
  • Direct Engineer Support
Refer & Earn

বন্ধুকে রেফার করুন, ফ্রি টোকেন নিন।

আপনার রেফারেল লিংক শেয়ার করুন। কেউ আপনার লিংক দিয়ে সাইনআপ করলেই দুজনেই $10 বোনাস টোকেন পাবেন।

🎁

$10 Bonus Tokens

আপনার রেফারেল লিংক দিয়ে কেউ সাইনআপ করলেই আপনি $10 বোনাস টোকেন পাবেন (claude-opus-4-8)।

💰

বন্ধুও একই পরিমাণ পাবে

আপনার রেফার করা বন্ধুও $10 বোনাস টোকেন পাবে। টোকেন কখনো expire হয় না — সাপ্তাহিক লিমিট শেষ হলে অটোমেটিক এখান থেকে খরচ হবে।

Model Catalog

Top Models On Your Key

One key, one endpoint — and Anthropic's top models run directly inside Claude Desktop. These six are the picks we would start you on, including the free-tier claude-opus-4-8. The full catalogue — every provider, every model — is one click away.

* Pricing is per million tokens and billed per request — no monthly minimum per model. The free weekly tier covers exactly one model — claude-opus-4-8, a genuine frontier model, not a stripped-down mini. Pro and Team unlock every model, higher rate limits and priority routing.

Build Patterns

What You Can Build

Common ways developers wire SleepyAI into a project — one key, no provider lock-in.

AI Research Assistant

A document Q&A tool that embeds your files, retrieves the relevant chunks and answers with claude-opus-5 — no DevOps, just HTTP calls.

Prototype on the free weekly claude-opus-4-8 allowance, then switch to claude-opus-5 by changing one model string.
claude-opus-4-8 → claude-opus-5
Free plan to prototype, Pro to scale

Automated Content Pipeline

Draft with a cheap model like gpt-5.2-mini, then refine the output with a thinking model like claude-opus-5-thinking — both through the same endpoint.

Mixing two providers in one pipeline needs one key and one invoice, not two billing accounts and two SDKs.
gpt-5.2-mini + claude-opus-5-thinking
Works from any language over plain HTTP

Customer Support Agent

Route routine tickets to a low-cost model such as deepseek-v4 and escalate the hard ones to claude-opus-5, deciding per request in your own code.

Cost routing is a model string in your request body — you are not tied to whichever provider you signed the contract with.
deepseek-v4 → claude-opus-5
One key, one invoice, per-request choice
FAQ

Frequently Asked Questions

Clear, zero-fluff answers about models, limits, pricing, and integration.

support@sleepyai.org

Start Building Today.

Connect straight to Claude Desktop and get instant API access to every frontier model — Claude, GPT, Gemini, Grok and DeepSeek. Start free on claude-opus-4-8. No credit card required. Up and running in minutes.

Direct Claude Desktop connection — key in, base URL set, done.
Free weekly limit on claude-opus-4-8 — no credit card needed.
Pro and Team unlock every model from every provider behind one key, plus the direct Claude Desktop connection.
Setup tutorials and developer support included.
Get Free API Access
← Back to home Get API Key
Catalog
Providers
All API Models

All Models

Sleepy Space
Pro | 12 Members
U
Your Zone
user@example.com
user@example.com
Claude Desktop-এ সরাসরি কানেক্ট করুন — key বসান, base URL https://sleepyai-api.vercel.app/v1 দিন, ব্যাস। Setup গাইড দেখুন ›

API Command Center

All Systems Operational
Real-time monitoring powered by AI • Last updated: just now
Global API Distribution
Request distribution across regions
15 Active Regions +5
Endpoint Distribution
Request distribution by endpoint
HTTP Methods
Request by method type
AI Insights
Intelligent recommendations
Active
Top Endpoints
Most requested APIs today
Active
Plans
Choose the right plan for your needs.
STARTER

Free Plan

ছোট প্রজেক্ট ও API পরীক্ষার জন্য পারফেক্ট।

৳0
  • একটি ফ্রন্টিয়ার মডেল ফ্রি — claude-opus-4-8 (200K context)
  • সাপ্তাহিক $10 টোকেন লিমিট
  • RPM 10
  • REST API অ্যাক্সেস
POPULAR
PRO

Pro Plan

যারা বেশি পাওয়ার আর হায়ার লিমিট চান তাদের জন্য।

৳299/মাস
  • Claude Desktop-এ সরাসরি কানেক্ট
  • সব মডেল অ্যাক্সেস
  • সাপ্তাহিক $250 টোকেন লিমিট
  • RPM 40
  • সরাসরি Native Provider Pipeline সংযোগ
  • Referral Bonus প্রোগ্রাম
  • Priority Support + Setup Tutorial
TEAM

Team Plan

প্রোডাকশন AI প্রোডাক্ট বানানো টিম ও কোম্পানির জন্য।

৳698/মাস
  • Claude Desktop-এ সরাসরি কানেক্ট
  • সব Pro ফিচার অন্তর্ভুক্ত
  • সাপ্তাহিক $600 টোকেন লিমিট
  • RPM 60
  • Shared Team API Key Management
  • Dedicated Onboarding Call
  • Direct Engineer Support
Referrals
Share your link — earn rewards when friends join SleepyAI.
YOUR REFERRAL LINK
🎁 আপনি পাবেন
প্রতি বন্ধুর জন্য $10 bonus token (claude-opus-4-8)
👥 আপনার বন্ধু পাবে
সাইনআপ করলেই $10 bonus token (claude-opus-4-8)

Total Referrals

0

Rewarded

0

Bonus Token Balance

0
Never expires — claude-opus-4-8 এ auto spend হয়

How it works

  1. আপনার referral link বন্ধুদের সাথে শেয়ার করুন।
  2. তারা আপনার link দিয়ে signup করলেই reward activate হবে।
  3. দুজনের জন্যই bonus token automatically credit হয়ে যাবে।
🏆 Referral History
🎁

এখনো কোনো referral নেই। Link শেয়ার করুন!

API Keys
Create, reveal, copy and revoke the keys that authenticate your requests.
Base endpoint ● operational
Use this base URL for all API requests.
https://sleepyai-api.vercel.app/v1
Your keys
You have 0/ keys. Keys are shown in full only once, right after creation.
Name Key Created Last used Status Actions

No API keys yet.
Create your first key to start using the API.

Using your key
Send the key as a Bearer token on every request. Never commit it to a public repo.
curl https://sleepyai-api.vercel.app/v1/chat/completions \
  -H "Authorization: Bearer $SLEEPY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"claude-sonnet-5","messages":[{"role":"user","content":"Hi"}]}'

API Endpoints

Manage and monitor your API endpoints • 8 total endpoints

Analytics

Live
Traffic, latency and error rates across every endpoint
Request Volume
Total requests over the selected range
Status Codes
Response distribution
Latency by Endpoint
p95 response time (ms)
Recent Requests
Live request log
MethodEndpointStatusLatencyTime

AI Insights

Active
Intelligent recommendations generated from your traffic patterns

Documentation

Guides, references and code samples for the SleepyAI API
Base URL
OpenAI-compatible — drop-in replacement base URL
Operational
https://sleepyai-api.vercel.app/v1
Usage
Monitor your API usage and token consumption.
0
Total Requests
0
Total Tokens
$0.0000
Total Cost (30d)
0
Cached Tokens

Weekly Allowance

$10.00
of $10.00 free weekly allowance remaining

Extra Credit Balance

$0.00
purchased credits — used after allowance exhausted
Free Model Tokens
claude-opus-4-8 — drawn from the weekly allowance
0 of 0.5M input tokens this week
≈ $10.00 at $20.00/M input — resets weekly, Monday 00:00 UTC
Bonus Token Balance
Referral-earned free-model tokens — never expire, spent on claude-opus-4-8 after your weekly allowance runs out.
0 tokens
Plan Limits — Free plan
5h Spend$0.0000
24h Spend$0.0000
Weekly Spend$0.0000 / $10.00
Requests per minute0 / 10
Transaction History
Extra Credit Balance: $0.00
Completed Transactions
Type Description Amount Status Date

No transactions yet.

Invoices
Your billing history and invoices.

No invoices yet.

Settings

Account, workspace and dashboard preferences
Account
How you appear across the workspace
Telegram community
Preferences
Theme and notification behaviour
Day mode
Switch the dashboard between light and dark. Same as the ☽/☀ button in the topbar.
AI insight alerts
Notify me when a new recommendation is generated.
Anomaly detection
Block suspicious request bursts automatically.
Weekly usage email
A summary of spend and requests every Monday.
More
Billing and plan pages

Plans

Compare Free, Pro and Team

Usage

Token spend and allowance

Transactions

Payment history

Invoices

Download receipts

Referrals

Invite and earn bonus tokens

Models

Every model and its pricing

Danger Zone
Deleting the account removes every key immediately. This cannot be undone.
Available Models
সব মডেল, একটাই key — claude-opus-4-8 ফ্রি (সাপ্তাহিক লিমিট), বাকিগুলো Pro ও Team প্ল্যানে। প্রতিটি মডেল Claude Desktop থেকেও সরাসরি ব্যবহার করা যায়।
Anthropic · Claude
claude-opus-4-8 Free · Pro · Team

আগের জেনারেশনের ফ্ল্যাগশিপ, এখনো জটিল reasoning ও coding-এ দুর্দান্ত — SleepyAI-এর একমাত্র free-tier মডেল, Free প্ল্যানে সাপ্তাহিক লিমিটের মধ্যে পাওয়া যায়

Input: $20.00/M tokens Output: $60.00/M tokens
✔ Available
claude-opus-4-8-thinking Pro · Team

Extended thinking সহ claude-opus-4-8

Input: $20.00/M tokens Output: $60.00/M tokens
✔ Available
claude-opus-5 Pro · Team

Anthropic-এর সবচেয়ে উন্নত মডেল

Input: $25.00/M tokens Output: $75.00/M tokens
✔ Available
claude-opus-5-thinking Pro · Team

Extended thinking সহ claude-opus-5

Input: $25.00/M tokens Output: $75.00/M tokens
✔ Available
claude-sonnet-5 Pro · Team

ব্যালান্সড workhorse — production traffic এর ডিফল্ট চয়েস

Input: $3.00/M tokens Output: $15.00/M tokens
✔ Available
claude-haiku-4-5 Pro · Team

সবচেয়ে দ্রুত ও সস্তা Claude — classification, routing, chat

Input: $1.00/M tokens Output: $5.00/M tokens
✔ Available
OpenAI · GPT
gpt-5.2 Pro · Team

OpenAI-এর ফ্ল্যাগশিপ — instruction following ও tool calling এ চমৎকার

Input: $12.00/M tokens Output: $36.00/M tokens
✔ Available
gpt-5.2-mini Pro · Team

সস্তা high-volume কাজের জন্য

Input: $0.45/M tokens Output: $1.80/M tokens
✔ Available
gpt-5-codex Pro · Team

Agentic coding ও বড় refactor এর জন্য tuned

Input: $10.00/M tokens Output: $30.00/M tokens
✔ Available
o4-mini Pro · Team

কম খরচে reasoning — math ও planning চেইনে ভালো

Input: $1.10/M tokens Output: $4.40/M tokens
✔ Available
gpt-4.1 Pro · Team

পুরনো integration এর জন্য stable legacy model

Input: $2.00/M tokens Output: $8.00/M tokens
✔ Available
Google · Gemini
gemini-3-pro Pro · Team

1M টোকেন context, শক্তিশালী multimodal

Input: $2.50/M tokens Output: $10.00/M tokens
✔ Available
gemini-3-flash Pro · Team

খুব দ্রুত ও খুব সস্তা — bulk document processing

Input: $0.30/M tokens Output: $1.20/M tokens
✔ Available
gemini-2-5-pro Pro · Team

2.5 behaviour এ pinned টিমের জন্য

Input: $1.25/M tokens Output: $5.00/M tokens
✔ Available
xAI · Grok
grok-4.1 Pro · Team

Live web search সহ শক্তিশালী reasoning

Input: $5.00/M tokens Output: $15.00/M tokens
✔ Available
grok-4-fast Pro · Team

কম latency, ফ্রেশ web context

Input: $0.50/M tokens Output: $1.50/M tokens
✔ Available
DeepSeek
deepseek-v4 Pro · Team

সেরা দাম-থেকে-কোয়ালিটি অনুপাত — bulk job এ সবচেয়ে জনপ্রিয়

Input: $0.28/M tokens Output: $1.10/M tokens
✔ Available
deepseek-r2 Pro · Team

কম খরচে লম্বা chain-of-thought reasoning

Input: $0.55/M tokens Output: $2.20/M tokens
✔ Available
Open Weight
llama-4-maverick Pro · Team

Meta open-weight — পরে self-host করা যায়

Input: $0.35/M tokens Output: $1.40/M tokens
✔ Available
mistral-large-3 Pro · Team

Multilingual, EU data residency অপশন

Input: $2.00/M tokens Output: $6.00/M tokens
✔ Available

Choose an endpoint

Pick an endpoint to load it into the tester with its default parameters.

Notifications

Alerts from your workspace.

Create API key

Give the key a name so you can recognise it later.

Log in

Welcome back. Log in to open your dashboard.
No account yet? Sign up
Documentation
Quickstart
Claude Desktop-এ সরাসরি কানেক্ট করুন — অথবা SleepyAI API দিয়ে মাত্র কয়েক মিনিটে শুরু করুন।

Claude Desktop-এর সাথে সরাসরি কানেকশনMAIN FOCUS

এটাই SleepyAI-এর প্রধান ব্যবহার। Claude Desktop খুলে Settings-এ যান, আপনার SleepyAI API key বসান, আর base URL দিন:

https://sleepyai-api.vercel.app/v1

Desktop app সরাসরি আমাদের API-এর সাথে কথা বলে, কারণ আমরা native Anthropic Messages API-compatible। এরপর Anthropic-এর top মডেলগুলো ডেস্কটপ অ্যাপ থেকেই সিলেক্ট করতে পারবেন — claude-opus-5, claude-opus-4-8, claude-sonnet-5, claude-haiku-4-5

Step 1 — API Key তৈরি করুন

Dashboard → API Keys → Generate New Key বাটন চাপুন। Key টি secure জায়গায় রাখুন।

Step 2 — Base URL

https://sleepyai-api.vercel.app/v1

এই endpoint একইসাথে Anthropic-compatible এবং OpenAI-compatible — সব provider এর সব model এখান থেকেই support করে, আর Claude Desktop এই একই URL দিয়েই সরাসরি চলে

Step 3 — প্রথম request পাঠান

curl https://sleepyai-api.vercel.app/v1/messages \
  -H "x-api-key: YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "claude-sonnet-5",
    "max_tokens": 1024,
    "messages": [{"role":"user","content":"Hello!"}]
  }'

Python দিয়ে

import anthropic

client = anthropic.Anthropic(
    api_key="YOUR_API_KEY",
    base_url="https://sleepyai-api.vercel.app/v1"
)

message = client.messages.create(
    model="claude-sonnet-5",
    max_tokens=1024,
    messages=[{"role": "user", "content": "Hello!"}]
)
print(message.content)
Claude Desktop — Direct ConnectionMAIN FOCUS
SleepyAI-এর প্রধান ব্যবহার: Claude Desktop অ্যাপ থেকে সরাসরি আমাদের API।
আমরা native Anthropic Messages API-compatible, তাই Claude Desktop আমাদের endpoint-কে নিজের provider হিসেবেই ব্যবহার করতে পারে। সেটআপে ~২ মিনিট লাগে।

Step 1 — API key নিন

Dashboard → API KeysGenerate New Key। secret একবারই দেখানো হয় — কপি করে রাখুন।

Step 2 — Base URL সেট করুন

Claude Desktop → Settings → API/provider সেকশনে base URL দিন:

https://sleepyai-api.vercel.app/v1

API key field-এ আপনার SleepyAI key বসান (header: x-api-key)।

Step 3 — মডেল বেছে নিন

যেকোনো Anthropic top মডেল desktop app থেকেই সিলেক্ট করা যাবে — claude-opus-5, claude-opus-4-8, claude-sonnet-5, claude-haiku-4-5। Claude Desktop-এর direct connection Pro ও Team প্ল্যানের ফিচার।

যাচাই করুন

curl https://sleepyai-api.vercel.app/v1/messages \
  -H "x-api-key: YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"claude-opus-4-8","max_tokens":64,
      "messages":[{"role":"user","content":"ping"}]}'

200 এলে Claude Desktop-ও একইভাবে কাজ করবে। 401 এলে key, 403 এলে plan (Free-তে শুধু claude-opus-4-8) চেক করুন।

Authentication
API key দিয়ে সব request authenticate করুন।

Header

প্রতিটি request এ এই header পাঠান:

x-api-key: sk-sleepy-xxxxxxxxxxxx

অথবা Anthropic SDK এর জন্য: Authorization: Bearer YOUR_KEY

Security Tips

• API key কখনো public repository তে push করবেন না
• Environment variable ব্যবহার করুন (.env file)
• Key compromise হলে Dashboard থেকে revoke করুন
• প্রতিটি project এর জন্য আলাদা key ব্যবহার করুন
Available Models
SleepyAI তে সব frontier মডেল উপলব্ধ — একটাই API key। Claude Desktop-এ সরাসরি Anthropic-এর top মডেলগুলো সিলেক্ট করা যায়। Free প্ল্যানে শুধু claude-opus-4-8 (সাপ্তাহিক লিমিটের মধ্যে) — কোনো কাটছাঁট করা mini মডেল নয়, একটা আসল frontier মডেল। বাকি মডেলগুলো শুধু Pro ও Team প্ল্যানে।
Anthropic · Claude
claude-opus-5 Pro · Team

সবচেয়ে শক্তিশালী মডেল। Agentic coding, deep research ও কঠিন কাজের জন্য।

Input: $25/M tokens Output: $75/M tokens Context: 300K
claude-opus-5-thinking Pro · Team

Extended thinking সহ claude-opus-5। thinking budget আপনি নিয়ন্ত্রণ করবেন।

Input: $25/M tokens Output: $75/M tokens Context: 300K
claude-sonnet-5 Pro · Team

ব্যালান্সড workhorse — production traffic ও high-volume agent এর ডিফল্ট চয়েস।

Input: $3/M tokens Output: $15/M tokens Context: 1M
claude-opus-4-8 Free · Pro · Team

আগের ফ্ল্যাগশিপ। জটিল reasoning ও coding এ এখনো দুর্দান্ত। SleepyAI-এর একমাত্র ফ্রি মডেল — Free প্ল্যানে সাপ্তাহিক লিমিটের মধ্যে ব্যবহার করা যায়।

Input: $20/M tokens Output: $60/M tokens Context: 200K
claude-opus-4-8-thinking Pro · Team

Extended thinking সহ claude-opus-4-8। গভীর multi-step reasoning এ অতুলনীয়।

Input: $20/M tokens Output: $60/M tokens Context: 200K
claude-haiku-4-5 Pro · Team

সবচেয়ে দ্রুত ও সস্তা Claude — classification, extraction ও routing এ আদর্শ।

Input: $1/M tokens Output: $5/M tokens Context: 200K
OpenAI · GPT
gpt-5.2 Pro · Team

OpenAI-এর ফ্ল্যাগশিপ। Instruction following, function calling ও JSON mode এ নির্ভরযোগ্য।

Input: $12/M tokens Output: $36/M tokens Context: 400K
gpt-5.2-mini Pro · Team

একই family, অনেক কম খরচ — chat frontend ও summarisation এ আদর্শ।

Input: $0.45/M tokens Output: $1.80/M tokens Context: 400K
gpt-5-codex Pro · Team

লম্বা refactor, repo-wide edit ও agent loop এর জন্য tuned।

Input: $10/M tokens Output: $30/M tokens Context: 256K
o4-mini Pro · Team

Reasoning-optimised ও সস্তা — math, planning ও evaluation chain এ ভালো।

Input: $1.10/M tokens Output: $4.40/M tokens Context: 200K
gpt-4.1 Pro · Team

পুরনো integration ও tuned prompt suite এর জন্য stable legacy।

Input: $2/M tokens Output: $8/M tokens Context: 1M
Google · Gemini
gemini-3-pro Pro · Team

পুরো codebase বা contract এক কলেই পাঠান। লম্বা context এ সেরা ভ্যালু।

Input: $2.50/M tokens Output: $10/M tokens Context: 1M
gemini-3-flash Pro · Team

বড় ডকুমেন্ট scale এ প্রসেস করার সবচেয়ে সস্তা উপায় — RAG ও bulk tagging।

Input: $0.30/M tokens Output: $1.20/M tokens Context: 1M
gemini-2-5-pro Pro · Team

2.5 behaviour এ pinned টিমের জন্য। একই multimodal input, একই লম্বা window।

Input: $1.25/M tokens Output: $5/M tokens Context: 1M
xAI · Grok
grok-4.1 Pro · Team

Live search সহ শক্তিশালী reasoning — news ও market প্রশ্নের জন্য।

Input: $5/M tokens Output: $15/M tokens Context: 256K
grok-4-fast Pro · Team

সস্তা Grok tier — chat product এ ফ্রেশ web context দরকার হলে।

Input: $0.50/M tokens Output: $1.50/M tokens Context: 256K
DeepSeek
deepseek-v4 Pro · Team

Frontier-এর কাছাকাছি coding ও math, দামের ভগ্নাংশে। Bulk job এ সবচেয়ে জনপ্রিয়।

Input: $0.28/M tokens Output: $1.10/M tokens Context: 128K
deepseek-r2 Pro · Team

কম খরচে লম্বা chain-of-thought — evaluation harness ও self-consistency voting এ।

Input: $0.55/M tokens Output: $2.20/M tokens Context: 128K
Open Weight
llama-4-maverick Pro · Team

Meta open-weight — এখানে prototype করুন, পরে একই prompt দিয়ে self-host করুন।

Input: $0.35/M tokens Output: $1.40/M tokens Context: 256K
mistral-large-3 Pro · Team

শক্তিশালী multilingual পারফরম্যান্স, EU data residency অপশন সহ।

Input: $2/M tokens Output: $6/M tokens Context: 128K
Completions API
Anthropic ও OpenAI — দুই ফরম্যাটেই compatible messages endpoint।

POST /messages

POST https://sleepyai-api.vercel.app/v1/messages

Request Body

{
  "model": "claude-opus-5",
  "max_tokens": 2048,
  "messages": [
    {"role": "user", "content": "আপনার message এখানে"}
  ],
  "system": "Optional system prompt"
}

Response

{
  "id": "msg_xxxx",
  "type": "message",
  "role": "assistant",
  "content": [{"type": "text", "text": "Response text"}],
  "model": "claude-opus-5",
  "usage": {"input_tokens": 10, "output_tokens": 50}
}
Rate Limits
প্রতিটি plan এর rate limit।
Plan RPM Weekly Limit Models
Free 10 $10 শুধু claude-opus-4-8 (1টি)
Pro 40 $250 সব মডেল
Team 60 $600 সব মডেল
Error Codes
Common API error codes এবং সমাধান।
401
Unauthorized

API key ভুল বা missing। x-api-key header check করুন।

429
Rate Limited

RPM limit পার হয়ে গেছে। কিছুক্ষণ অপেক্ষা করুন বা Pro plan এ upgrade করুন।

402
Credit Exhausted

Weekly বা monthly limit শেষ। Dashboard থেকে top-up করুন বা পরের সপ্তাহ পর্যন্ত অপেক্ষা করুন।

403
Model Access Denied

এই model আপনার plan এ নেই। Free plan এ শুধু claude-opus-4-8 চলে; বাকি model গুলো (claude-opus-5, claude-sonnet-5, gpt-5.2, gemini-3-pro, grok-4.1 সহ) ব্যবহার করতে Pro বা Team plan দরকার।