Recommended models
Find the recommended built-in models for GPT for Sheets and GPT for Excel, and choose the one that best fits your use case and budget. For more information about the full range of supported models, see AI providers and models supported by GPT for Work.
Agent models​
| Best | Balanced | Cheapest |
|---|---|---|
Fable 5.1 (Medium) | GPT-5.6 Luna (High) | GPT-5.4 (Low) |
For the benchmark results behind these recommendations, see Agent model benchmark.
Examples​
-
For simple or straightforward tasks such as creating formulas or fixing errors, use GPT-5.6 Luna at High reasoning effort.
-
For high-level tasks such as financial modeling, overall formatting, or visual tasks, use Fable 5.1 at Medium reasoning effort.
Bulk models​
The same bulk models are recommended whether the Agent delegates a bulk task, or you run bulk tools or GPT functions yourself.
| Use case | Best | Balanced | Cheapest |
|---|---|---|---|
Content generation, translation, categorization, extraction, scoring | gpt-5.6-sol | gpt-5.4-mini | gpt-5.4-nano |
Enrichment from the web | gpt-5-search-api | perplexity-low | perplexity-fast |
Agent model benchmark​
To compare Agent models, we run an internal benchmark of roughly 200 everyday spreadsheet tasks, such as writing a formula, inserting a column, creating a chart, or formatting a header. Each model runs every task once with the same spreadsheets and prompts, and each result is scored as a pass or a fail. Newly released models are always evaluated in real-world use before we recommend them, even when they score well in the benchmark.
The latest run is from September 2026, in which six models competed in GPT for Excel:
-
Fable 5.1 at Medium reasoning effort achieved the best overall performance and ranked first or tied for first in half of the task categories.
-
GPT-5.6 Luna at High reasoning effort offered the best balance of score and cost. It scored close to Fable 5.1 while consuming about half the credits per task.
The benchmark ran in GPT for Excel only. The Agent uses different tools in Google Sheets, so results in Sheets may differ.
Overall scores​
The following figure shows the share of tasks each model completed successfully.
| Model | Reasoning effort | Score |
|---|---|---|
| Fable 5.1 | Medium | 90.9% |
| GPT-6 Astra | Low | 89.8% |
| GPT-5.6 Luna | High | 86.6% |
| GPT-5.6 Sol | Medium | 86.6% |
| Opus 4.8 | High | 85.6% |
| GPT-5.4 | Low | 79.7% |
Scores vs. credit consumption​
The following figure shows each model's score against its credit consumption per task, relative to GPT-5.6 Luna at High reasoning effort. Credit consumption covers the Agent model only, not the bulk models the Agent delegates to, and assumes built-in models on a subscription. For example, Fable 5.1 consumes about twice as many credits per task as GPT-5.6 Luna. The dashed line connects the models that offer the best trade-off: Every model off the line is matched or outscored by a model that consumes fewer credits.
| Model | Reasoning effort | Score | Credit consumption relative to GPT-5.6 Luna (High) |
|---|---|---|---|
| Fable 5.1 | Medium | 90.9% | 2.14× |
| GPT-6 Astra | Low | 89.8% | 2.15× |
| GPT-5.6 Luna | High | 86.6% | 1× |
| GPT-5.6 Sol | Medium | 86.6% | 1.85× |
| Opus 4.8 | High | 85.6% | 1.79× |
| GPT-5.4 | Low | 79.7% | 0.67× |
Scores vs. speed​
The following figure shows each model's score against the average time it took to complete a task. Shorter times are better. The dashed line connects the models that offer the best trade-off: Every model off the line is matched or outscored by a faster model.
| Model | Reasoning effort | Score | Average time per task |
|---|---|---|---|
| Fable 5.1 | Medium | 90.9% | 50s |
| GPT-6 Astra | Low | 89.8% | 47s |
| GPT-5.6 Luna | High | 86.6% | 43s |
| GPT-5.6 Sol | Medium | 86.6% | 46s |
| Opus 4.8 | High | 85.6% | 44s |
| GPT-5.4 | Low | 79.7% | 35s |
Rankings by task category​
The following table shows how the models rank in each task category. Models with the same score in a category share a rank.
| Task category | Fable 5.1Medium | GPT-6 AstraLow | GPT-5.6 LunaHigh | GPT-5.6 SolMedium | Opus 4.8High | GPT-5.4Low |
|---|---|---|---|---|---|---|
| Bulk content generation | 1st | 2nd | 1st | 2nd | 4th | 3rd |
| Bulk data processing | 2nd | 2nd | 1st | 2nd | 2nd | 3rd |
| Charts | 2nd | 1st | 4th | 4th | 3rd | 5th |
| Data operations | 1st | 2nd | 3rd | 2nd | 2nd | 4th |
| Formatting | 1st | 1st | 3rd | 3rd | 2nd | 2nd |
| Formulas | 1st | 2nd | 1st | 1st | 2nd | 3rd |
| Pivot tables | 2nd | 2nd | 1st | 1st | 2nd | 2nd |
| Spreadsheet manipulation | 2nd | 1st | 3rd | 2nd | 2nd | 4th |