Skip to main content
Chibi supports a wide range of AI providers. You can configure which models to use for text, images, and voice, as well as fine-tune their behavior.

Master API Keys

To use a specific provider, you must set its corresponding API key in your .env file.

Model Configuration

You can specify which models Chibi should use by default.

Available Models (Examples)

Please note that the full list of supported models is much larger, it is enormous. Here only few examples are provided:
  • OpenAI: gpt-5.2, gpt-5.1, o3, o4-mini
  • Anthropic: claude-sonnet-4-5-20250929, claude-haiku-4-5-20251001
  • Grok: grok-4-1-fast-reasoning, grok-4-1-reasoning, grok-beta
  • Gemini: gemini-2.5-pro, gemini-3-pro
  • Alibaba: qwen3-max, qwen-max
  • MiniMax: MiniMax-M2.5, MiniMax-M2.5-highspeed; Image-01 (image)
  • ZhipuAI: glm-5, glm-4-flash, glm-4, glm-4-vision

Text Generation Parameters

Fine-tune how the LLM generates text.

Image Generation Settings

Configure the quality and dimensions of generated images.

Provider-Specific Image Sizes

Some providers require specific resolutions. You can override the default IMAGE_SIZE for them:
  • IMAGE_SIZE_NANO_BANANA (for Gemini Image models)
  • IMAGE_SIZE_IMAGEN (for Imagen 4.0)
  • IMAGE_SIZE_ALIBABA (for Wan/Qwen Image)

Voice & Audio (STT/TTS)

Configure Speech-to-Text (STT) and Text-to-Speech (TTS) capabilities.

MiniMax TTS Specific Settings

For MiniMax Text-to-Speech, you can configure the following: