Models can misreport their name and version at runtime. This is because they are often trained on data generated by previous versions of themself or other models.
A context window is the maximum span of tokens (text and code) a model can consider at once. The more prompts, files, and responses in a session, the more context is consumed.Firebender intelligently summarizes and shifts context around to balance speed and accuracy.An estimate of tokens used is provided:
Companies may not want to support certain models or providers, and can restrict what models their team has access to. Team admins will need to add later models to the list when new models are released by providers.Get started: Model restrictions
Firebender plans include usage at model provider API rates. For example, $30 of included usage on the Developer plan will be consumed based on your model selection and its price.Usage limits are shown in editor based on your current consumption. All prices are per million tokens.