/00 — boot sequence

Hello.

/01 — Free credits

Free credits & free tiers

Where to get real inference without reaching for a card.

▸Google AI Studio

Gemini 2.5 Flash/Pro, Gemini 3.5 Flash — 1M context, 1500 req/day free tier, no credit card required

free apino cardBest free-to-good ratio. Gemini 2.5 Flash is competitive with GPT-4o-mini.
▸Groq Cloud

Llama, Mixtral, Gemma models at record inference speed. Rate-limited free tier, no credit card needed

free apino cardFastest tokens/sec for $0. Great for prototyping and real-time apps.
▸OpenRouter

20 RPM free tier on select models (Llama 3, Gemma, Qwen, Nemotron). Up to 1000 req/day with $10 lifetime topup

free apino cardFilter :free models. One API key, 20+ free models.
▸GitHub Models

GPT-4o, Claude Sonnet, Llama 3, Mistral Large inference inside your GitHub account. Free with GitHub PAT

free apino cardGreat for dev + CI/CD prototyping under your existing GitHub account.
▸Hugging Face Inference

Serverless Inference API with generous free daily quota. Thousands of community models + Spaces hosting

free apino cardEnormous model catalog. Best-effort latency, ideal for experimentation.
▸Mistral AI

Mistral Large 3, Medium 3.5, Small 4, Codestral — ~1B tokens/month free tier, no credit card

free apino cardBest free European provider. Codestral is excellent for code generation.
▸NVIDIA NIM

Llama 3, Nemotron, Mistral, and more on NVIDIA's optimized inference stack. Phone verification required

free apiNVIDIA-optimized throughput. Requires phone, no credit card.
▸OpenCode Zen

Free inference for Hermes models, DeepSeek, Qwen, and more. OpenAI-compatible API

free apino cardZero-cost inference for open models. Powers this very portfolio.
▸Cloudflare Workers AI

Llama, Mistral, Qwen models via Cloudflare's global edge network. 100K requests/day free

free apino cardEdge-deployed AI with global latency. Generous free tier, no card needed.
▸Cohere

Command A, Command R+ models — 1000 API calls/month free tier, non-commercial use. No credit card

free apino cardStrong for RAG and enterprise search use cases with 256K context.
▸Vercel AI SDK (AI SDK)

AI SDK with built-in support for OpenAI, Anthropic, Google, Mistral. Vercel AI Gateway free tier included

free apino cardFramework + gateway in one. Great for Next.js AI features.
▸Cerebras

Giga-scale AI inference on wafer-scale chips. Llama and Mistral models at extreme speeds, free tier available

free apino cardFastest hardware inference. Wafer-scale, extreme throughput.
▸Fireworks AI

Llama, Mixtral, DeepSeek models with high throughput. $1 in trial credits, no credit card for initial signup

no cardHigh-throughput inference with competitive pricing after trial.
▸SambaNova Cloud

Llama 3, DeepSeek, Qwen models on SN40L chips. Free tier with rate limits, no credit card

free apino cardSN40L RDU-powered. Competitive speed for open models.