Inference usage is deducted from your available credits. New accounts receive
a one-time $6 welcome credit, valid for one calendar year from issuance.
See Pricing for rates and how to add credits.
Data security
Read about zero data retention for inference. Token Factory retains operational metadata; the Data Security page explains what is retained and how volatile prompt caching works.API compatibility
OpenAI-compatible API
Use Chat Completions, the Responses API, model listing, and supported
OpenAI SDK features.
Anthropic-compatible API
Use the Messages API, token counting, the Anthropic SDK, and Claude Code.
Models and pricing
Compare the models currently served by Token Factory, including context
windows, capabilities, and per-token rates.
Usage API
Query account-scoped usage from
GET /api/v1/usage with your API key.Start here
Send your first request
Create an API key and call
POST /v1/chat/completions.Protect your API key
Choose the correct authentication header and store keys safely.
Connect a client or tool
Configure an SDK or coding agent.
Read the API reference
Review endpoint schemas generated from the public OpenAPI 3.1 specification.
Community and support
Join the Discord
Ask questions in #get-help and share feedback in #feedback. No invitation
required.
Contact support
Use the dashboard support form or email for account and billing questions.