Model capabilities
Reviewed provider facts, with Pumpkin request policy shown separately. Unknown does not mean unsupported or unlimited.
Model pricing · Editor seat pricing · Catalog JSON
Skill ratings and task descriptions are Pumpkin editorial judgments, not provider benchmarks. Fireworks models are opt-in; a catalog row is not a guarantee of account or serverless availability. Local models remain managed in the editor.
Reviewed .
Catalog revision
ebfc771294f78902ba21009cf9f4792859e5bf9c8cc0bcd56db795618ef348d3Fireworks live discovery: available; last successful fetch 2026-09-23T09:56:25Z.
| Model | Context and tools | Provider image input | Pumpkin client policy | Evidence and qualifications |
|---|---|---|---|---|
Claude Fable 5.1anthropic · defaultclaude-fable-5-1Editorial skill: 10 / 10 |
Context: 1000000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 600 Formats and limitsProvider formats: image/jpeg, image/png, image/gif, image/webp Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: 8000 / 8000 Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified. Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images. |
Long edge: 2576 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Hardest work: architecture, multi-service changes, long-context debugging, anything where a wrong call is expensive. Standard Messages output limit; thinking shares output budget. Adaptive thinking is always on; forced tool choice is unsupported. Verified Sep 22, 2026 https://platform.claude.com/docs/en/models/fable-5-1/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing |
Claude Opus 5.5anthropic · defaultclaude-opus-5-5Editorial skill: 9 / 10 |
Context: 1000000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 600 Formats and limitsProvider formats: image/jpeg, image/png, image/gif, image/webp Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: 8000 / 8000 Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified. Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images. |
Long edge: 2576 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Long-running agentic coding and knowledge work. Newer and cheaper than Opus 5, still below Fable. Standard Messages output limit; thinking shares output budget. Adaptive thinking is always on; forced tool choice is unsupported. Default effort is medium. Cache read is 0.05x base input, not the usual 0.1x. Verified Sep 22, 2026 https://platform.claude.com/docs/en/models/opus-5-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing |
Claude Opus 5anthropic · defaultclaude-opus-5Editorial skill: 8 / 10 |
Context: 1000000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 600 Formats and limitsProvider formats: image/jpeg, image/png, image/gif, image/webp Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: 8000 / 8000 Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified. Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images. |
Long edge: 2576 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Strong implementation and review when Fable is overkill: scoped features, correctness-sensitive edits, production-ready diffs. Standard Messages output limit; thinking shares output budget. Verified Sep 22, 2026 https://platform.claude.com/docs/en/models/opus-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing |
Claude Sonnet 5anthropic · defaultclaude-sonnet-5Editorial skill: 6 / 10 |
Context: 1000000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 600 Formats and limitsProvider formats: image/jpeg, image/png, image/gif, image/webp Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: 8000 / 8000 Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified. Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images. |
Long edge: 2576 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Everyday coding: routine features, refactors with a known shape, tests, docs. Mid-tier default for Anthropic. Standard Messages output limit; thinking shares output budget. Verified Sep 22, 2026 https://platform.claude.com/docs/en/models/sonnet-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing |
Claude Haiku 4.5anthropic · defaultclaude-haiku-4-5Editorial skill: 4 / 10 |
Context: 200000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 64000 Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 100 Formats and limitsProvider formats: image/jpeg, image/png, image/gif, image/webp Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: 8000 / 8000 Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified. Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images. |
Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Cheap, short jobs: mechanical edits, triage, titles, compaction, apply this exact change. Not for design or long hunts. Standard Messages output limit; thinking shares output budget. Verified Sep 22, 2026 https://platform.claude.com/docs/en/models/haiku-4-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing |
Grok 4.7xai · defaultgrok-4.7Editorial skill: 7 / 10 |
Context: 500000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: low, medium, high, xhigh |
Supported: Yes Count: Unknown Formats and limitsProvider formats: image/jpeg, image/png Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown 20 MiB per image; decoded versus encoded accounting unspecified. No image-count limit stated, but context limits still apply. Hard pixel dimensions and numeric provider resize recommendation unknown. |
Output request cap: 128000 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: xAI flagship for coding and agent work. Same price as 4.6, stronger on tools and harder implementation. Provider documents no separate text output limit. maxOutput=128000 is Pumpkin request policy, not a provider maximum. Reasoning cannot be disabled. Verified Sep 22, 2026 https://docs.x.ai/developers/models/grok-4.7https://docs.x.ai/developers/model-capabilities/images/understandinghttps://docs.x.ai/developers/pricing |
Grok 4.6xai · defaultgrok-4.6Editorial skill: 6 / 10 |
Context: 500000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: low, medium, high, xhigh |
Supported: Yes Count: Unknown Formats and limitsProvider formats: image/jpeg, image/png Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown 20 MiB per image; decoded versus encoded accounting unspecified. No image-count limit stated, but context limits still apply. Hard pixel dimensions and numeric provider resize recommendation unknown. |
Output request cap: 128000 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Fast workhorse: investigation, evidence collection, intermediate code, image analysis. Good value; less stubborn than Fable on long stuck turns. Provider documents no separate text output limit. maxOutput=128000 is Pumpkin request policy, not a provider maximum. Reasoning cannot be disabled. Verified Sep 22, 2026 https://docs.x.ai/developers/models/grok-4.6https://docs.x.ai/developers/model-capabilities/images/understandinghttps://docs.x.ai/developers/pricing |
GPT-5.6 Solopenai · defaultgpt-5.6-solEditorial skill: 5 / 10 |
Context: 1050000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Yes Effort values: none, low, medium, high, xhigh, max |
Supported: Yes Count: 1500 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified. Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity. |
Long edge: 2048 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Solid mid OpenAI: focused features, bug fixes with a known cause, normal coding loops. Maximum input 922000 tokens; reasoning and visible output share the output budget. Promotional prices at least through 2026-11-21; subsequent rates unknown. Verified Sep 22, 2026 https://developers.openai.com/api/docs/models/gpt-5.6-solhttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing |
GPT-5.6 Terraopenai · defaultgpt-5.6-terraEditorial skill: 4 / 10 |
Context: 1050000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Yes Effort values: none, low, medium, high, xhigh, max |
Supported: Yes Count: 1500 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified. Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity. |
Long edge: 2048 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Small, well-specified work: one-file edits, straightforward tests, low-stakes helpers. Maximum input 922000 tokens; reasoning and visible output share the output budget. Verified Sep 22, 2026 https://developers.openai.com/api/docs/models/gpt-5.6-terrahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing |
GPT-5.6 Lunaopenai · defaultgpt-5.6-lunaEditorial skill: 2 / 10 |
Context: 1050000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Yes Effort values: none, low, medium, high, xhigh, max |
Supported: Yes Count: 1500 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified. Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity. |
Long edge: 2048 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Cheapest OpenAI: tiny mechanical tasks, classification, short rewrites. Skip anything that needs judgment. Maximum input 922000 tokens; reasoning and visible output share the output budget. Verified Sep 22, 2026 https://developers.openai.com/api/docs/models/gpt-5.6-lunahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing |
GPT-6 Astraopenai · defaultgpt-6-astraEditorial skill: 10 / 10 |
Context: 1050000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Yes Effort values: low, medium, high, xhigh, max |
Supported: Yes Count: 1500 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified. Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity. |
Long edge: 65535 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: OpenAI flagship: hard implementation and long context, same class as Fable. Use when the problem is large or ambiguous. Maximum input 922000 tokens; reasoning and visible output share the output budget. Tool calling requires Responses; Chat Completions supports text only. Verified Sep 22, 2026 https://developers.openai.com/api/docs/models/gpt-6-astrahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing |
Muse Spark 1.3meta · defaultmuse-spark-1.3Editorial skill: 6 / 10 |
Context: 1048576 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: Unknown separate numeric maximum Pumpkin output request cap: 65536 tokens, not a documented provider maximum. Streaming: Yes Effort values: minimal, low, medium, high, xhigh, max |
Supported: Yes Count: 20 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Images as base64 data URLs or Files API references; audio understanding is degraded on 1.3 (use 1.2), and GIF uses the first frame. Meta documents no per-image or per-request image byte limits; Pumpkin applies its own conservative client policy rather than trusting an unspecified vendor cap. |
Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Meta's agentic reasoning model: multi-step tool loops and coding, 1M context, cheap cached input. Reasoning cannot be turned off (effort "none" returns HTTP 400); reasoning shares the output budget with visible text. Meta documents no per-request output ceiling; 65536 is a safe request cap below the 1,048,576 context window (input and output share one budget). Verified Sep 22, 2026 https://dev.meta.ai/docs/modelshttps://dev.meta.ai/docs/pricing-rate-limitshttps://dev.meta.ai/docs/reasoninghttps://dev.meta.ai/docs/image-understanding |
DeepSeek V4 Pro 0813fireworks · opt-inaccounts/fireworks/models/deepseek-v4-pro-0813Editorial skill: 6 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: Strongest DeepSeek: hard coding, agents and long reasoning; Fireworks' top open pick. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://app.fireworks.ai/models/fireworks/deepseek-v4-pro-0813https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Kimi K3fireworks · opt-inaccounts/fireworks/models/kimi-k3Editorial skill: 6 / 10Availability: available |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Kimi K3 flagship: hard coding and research agents, image input, 1M context; the priciest here. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/kimi-k3https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
GLM 5.3fireworks · opt-inaccounts/fireworks/models/glm-5p3Editorial skill: 6 / 10Availability: available |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: Newest GLM flagship: coding, reasoning and agent tool use; 1M context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/glm-5p3https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
GLM 5.2fireworks · opt-inaccounts/fireworks/models/glm-5p2Editorial skill: 5 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: GLM 5.2 flagship: coding and agent tool use; 1M context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/glm-5p2https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Kimi K2.7 Codefireworks · opt-inaccounts/fireworks/models/kimi-k2p7-codeEditorial skill: 6 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 262144 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Kimi K2.7 Code: coding-tuned agent model with image input; 262K context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/kimi-k2p7-codehttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Kimi K2.6fireworks · opt-inaccounts/fireworks/models/kimi-k2p6Editorial skill: 5 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 262144 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Kimi K2.6: agents with tool use, images and documents; 262K context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/kimi-k2p6https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Qwen 3.8 2.4T A95Bfireworks · opt-inaccounts/fireworks/models/qwen3p8-2p4t-a95bEditorial skill: 5 / 10Availability: unavailable |
Context: 262144 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: Qwen 3.8 (2.4T MoE, 95B active): strong general coding and reasoning; 262K context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Public model page says serverless not supported. No substitution of qwen3p8-max pricing. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/qwen3p8-2p4t-a95bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
MiniMax M3fireworks · opt-inaccounts/fireworks/models/minimax-m3Editorial skill: 5 / 10Availability: available |
Context: 512000 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: MiniMax M3: fast coding and agent model, good value; 512K context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Exact serving-model feature metadata says no images despite generic multimodal architecture prose. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/minimax-m3https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
DeepSeek V4.1 Flashfireworks · opt-inaccounts/fireworks/models/deepseek-v4p1-flashEditorial skill: 4 / 10Availability: available |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: DeepSeek V4.1 Flash with image input: quick coding and agent steps at low cost; 1M context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://app.fireworks.ai/models/fireworks/deepseek-v4p1-flashhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Nemotron 3 Ultra NVFP4fireworks · opt-inaccounts/fireworks/models/nemotron-3-ultra-nvfp4Editorial skill: 4 / 10Availability: available |
Context: 262144 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: NVIDIA Nemotron 3 Ultra (preview): large reasoning model; 262K context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/nemotron-3-ultra-nvfp4https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
DeepSeek V4 Flash 0731fireworks · opt-inaccounts/fireworks/models/deepseek-v4-flash-0731Editorial skill: 3 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: Fast, cheap DeepSeek: extraction, summaries and everyday agent steps; 1M context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://app.fireworks.ai/models/fireworks/deepseek-v4-flash-0731https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
DeepSeek V4 Flash Vision Experimentalfireworks · opt-inaccounts/fireworks/models/deepseek-v4-flash-vision-expEditorial skill: 3 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: DeepSeek V4 Flash with experimental image input; cheap, 1M context. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://app.fireworks.ai/models/fireworks/deepseek-v4-flash-vision-exphttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
GPT OSS 120Bfireworks · opt-inaccounts/fireworks/models/gpt-oss-120bEditorial skill: 3 / 10Availability: available |
Context: 131072 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: OpenAI open-weight 120B: solid reasoning and tool use at low cost; no images. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/gpt-oss-120bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
GLM 5.3 Flashfireworks · opt-inaccounts/fireworks/models/glm-5p3-flashEditorial skill: 3 / 10Availability: available |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Small, cheap GLM with image input: quick edits and lookups. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/glm-5p3-flashhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Muse Glimmer 30Bfireworks · opt-inaccounts/fireworks/models/muse-glimmer-30bEditorial skill: 3 / 10Availability: availableServerless retirement: Sep 25, 2026 |
Context: 131072 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Small 30B always-on agent model with image input: cheap and quick for simple tasks. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/muse-glimmer-30bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Nemotron Lightning 3.5 30B A3Bfireworks · opt-inaccounts/fireworks/models/nemotron-lightning-3p5-30b-a3bEditorial skill: 2 / 10Availability: available |
Context: 262144 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: No Discovery reports image input unsupported; curated support is conservatively disabled. |
Output request cap: 16384 tokens (also bounded by context). Image policy: Not applicable / Unknown. |
Details and sourcesEditorial guidance: Cheapest fast agent model (30B, 3B active): small edits, descriptions, triage. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/nemotron-lightning-3p5-30b-a3bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
Inklingfireworks · opt-inaccounts/fireworks/models/inklingEditorial skill: UnknownAvailability: available |
Context: 1048576 tokens Tools: Yes Reasoning: Unknown Output and reasoning detailsProvider output: Unknown separate numeric maximum Streaming: Unknown Effort values: Unknown |
Supported: Yes Count: 30 Formats and limitsProvider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds. No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy. |
Output request cap: 16384 tokens (also bounded by context). Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Fireworks serverless model with image input and a 1M context; capability not rated. Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning. Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits. Verified Sep 22, 2026 https://fireworks.ai/models/fireworks/inklinghttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models |
GPT-6 Solopenai · defaultgpt-6-solEditorial skill: 6 / 10 |
Context: 1050000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Yes Effort values: none, low, medium, high, xhigh, max |
Supported: Yes Count: 1500 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown GIF must be nonanimated. The image guide allows up to 512 MB total request payload; exact byte accounting is unspecified. Model-specific image dimensions remain Unknown. |
Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Complex coding and agent work when Astra is more than you need. Cheaper than GPT-5.6 Sol. Use Responses for built-in tools and function calling. Chat Completions function calling requires reasoning_effort=none. Exact model-specific image resize limits are not stated in the reviewed image guide; Pumpkin uses its conservative client policy, not an inferred provider maximum. Verified Sep 22, 2026 https://developers.openai.com/api/docs/models/gpt-6-solhttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing |
GPT-6 Lunaopenai · defaultgpt-6-lunaEditorial skill: 3 / 10 |
Context: 1050000 tokens Tools: Yes Reasoning: Yes Output and reasoning detailsProvider output: 128000 Streaming: Yes Effort values: none, low, medium, high, xhigh, max |
Supported: Yes Count: 1500 Formats and limitsProvider formats: image/png, image/jpeg, image/webp, image/gif Bytes per image: Unknown Aggregate encoded bytes: Unknown Hard width / height: Unknown / Unknown GIF must be nonanimated. The image guide allows up to 512 MB total request payload; exact byte accounting is unspecified. Model-specific image dimensions remain Unknown. |
Long edge: 1568 px Pumpkin image limitsPumpkin-selected image limits, not provider maxima:
Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly. |
Details and sourcesEditorial guidance: Cheapest OpenAI: high-volume tasks, classification, short rewrites. Not for design or long hunts. Use Responses for built-in tools and function calling. Chat Completions function calling requires reasoning_effort=none. Exact model-specific image resize limits are not stated in the reviewed image guide; Pumpkin uses its conservative client policy, not an inferred provider maximum. Verified Sep 22, 2026 https://developers.openai.com/api/docs/models/gpt-6-lunahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing |