Free models on Tokenator
These models are free to use — requests to them do not spend tokens from your bundle. Each one has its own daily allowance that renews every day.
This model is available on a free plan with set usage limits. Gemini 3.7 Flash is a multimodal model from Google designed for fast workflows, programming, and complex multi-step reasoning. It combines high speed with robust solutions for problems requiring sequential analysis and multiple steps.
This model is available on a free plan with set usage limits. Claude Opus 5 is Anthropic's flagship model for complex logical reasoning, programming, and long-term agent management. It is particularly effective in complex software development, code review and error detection, document and diagram analysis, complex office tasks, and coordinating multiple concurrent subagents. Claude Opus 5 reliably follows instructions and uses tools efficiently, even in lengthy, multi-step tasks. It is also well-suited for scenarios where low latency and more efficient token usage are essential.
This model is available on a free plan with pre-defined usage limits. MiniMax-M3 is a multimodal baseline model from MiniMax, focused on long-term agent work, programming, and tool use. It supports text, image, and video input, text output, and a context window of up to 1 million tokens. The model is built on the MiniMax Sparse Attention (MSA) mechanism, which uses key-value block selection instead of full attention. This significantly reduces computational overhead when working with large contexts — approximately 20x compared to the previous generation for a context of 1 million tokens — while also accelerating pre-population and decoding. MiniMax-M3 was initially trained as a multimodal model on interleaved data and is further optimized for long-term multi-step interactions. This makes it particularly well-suited for complex agent-based scenarios that require sequential execution of a large number of actions, tool use, and maintaining context throughout the task.
This model is available on a free plan with set usage limits. Grok 4.6 is SpaceXAI's most intelligent model, demonstrating cutting-edge results in programming, intellectual work, and STEM education.
4 free models, 4 available right now, they do not spend bundle tokens, daily allowance up to 4M tokens
Daily allowances of the free models
When a model runs out of its daily allowance it stops answering until the allowance renews, and your paid tokens stay untouched. The other free models keep working — each one has its own allowance.
| Model | API ID | Vendor | Context | Daily allowance |
|---|---|---|---|---|
| Free Grok 4.6 | free-grok-4.6 | xAI | 500K tokens | 4M tokens and 60 requests per day |
| Free MiniMax M3 | free-minimax-m3 | MiniMax | 1M tokens | 4M tokens and 60 requests per day |
| Free Gemini 3.7 Flash | free-gemini-3.7-flash | Gemini | 1M tokens | 4M tokens and 60 requests per day |
| Free Claude Opus 5 | free-claude-opus-5 | Anthropic | 1M tokens | 4M tokens and 60 requests per day |