mirror of
https://github.com/tiennm99/DocsGPT.git
synced 2026-10-04 10:13:06 +00:00
Rename the unused *_cost_per_token capability fields to USD per 1M tokens, add prompt-cache read/write rates, and ship list prices for the hosted catalogs. The old per-token keys still load, scaled, with a warning. docsgpt/pricing.py turns a call's token bins into a USD cost. Models with no declared rate cost $0 unless QUOTA_UNPRICED_RATE_PER_MILLION is set.
25 lines
911 B
YAML
25 lines
911 B
YAML
provider: google
|
|
defaults:
|
|
supports_tools: true
|
|
supports_structured_output: true
|
|
attachments: [pdf, image]
|
|
context_window: 1048576
|
|
|
|
models:
|
|
- id: gemini-3.1-pro-preview
|
|
display_name: Gemini 3.1 Pro (preview)
|
|
description: Most capable Gemini 3 model with advanced reasoning and agentic coding (preview)
|
|
# Priced at the >200k-token tier; long prompts are common with attachments.
|
|
input_cost_per_million: 4.0
|
|
output_cost_per_million: 18.0
|
|
- id: gemini-3.5-flash
|
|
display_name: Gemini 3.5 Flash
|
|
description: Frontier-class Flash for sustained performance on agentic and coding tasks
|
|
input_cost_per_million: 1.5
|
|
output_cost_per_million: 9
|
|
- id: gemini-3.1-flash-lite
|
|
display_name: Gemini 3.1 Flash-Lite
|
|
description: Cost-efficient frontier-class multimodal model for high-throughput workloads
|
|
input_cost_per_million: 0.25
|
|
output_cost_per_million: 1.5
|