Model capabilities

Reviewed provider facts, with Pumpkin request policy shown separately. Unknown does not mean unsupported or unlimited.

Model pricing · Editor seat pricing · Catalog JSON

Skill ratings and task descriptions are Pumpkin editorial judgments, not provider benchmarks. Fireworks models are opt-in; a catalog row is not a guarantee of account or serverless availability. Local models remain managed in the editor.

Reviewed .

Catalog revisionebfc771294f78902ba21009cf9f4792859e5bf9c8cc0bcd56db795618ef348d3

Fireworks live discovery: available; last successful fetch 2026-09-23T09:56:25Z.

Provider capabilities and Pumpkin-selected limits. Originals are preserved; processing applies to request copies.
ModelContext and toolsProvider image inputPumpkin client policyEvidence and qualifications
Claude Fable 5.1anthropic · defaultclaude-fable-5-1Editorial skill: 10 / 10

Context: 1000000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 600

Formats and limits

Provider formats: image/jpeg, image/png, image/gif, image/webp

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: 8000 / 8000

Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified.

Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images.

Long edge: 2576 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 3750656
  • Decoded allocation: 100000000 bytes
  • File bytes per image: 7000000 (Pumpkin cap; 7,000,000 file bytes encode to at most 9,333,336 base64 bytes, below the documented 10 MB)
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 28 px patches, budget 4784 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Hardest work: architecture, multi-service changes, long-context debugging, anything where a wrong call is expensive.

Standard Messages output limit; thinking shares output budget.

Adaptive thinking is always on; forced tool choice is unsupported.

Verified Sep 22, 2026

https://platform.claude.com/docs/en/models/fable-5-1/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing
Claude Opus 5.5anthropic · defaultclaude-opus-5-5Editorial skill: 9 / 10

Context: 1000000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 600

Formats and limits

Provider formats: image/jpeg, image/png, image/gif, image/webp

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: 8000 / 8000

Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified.

Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images.

Long edge: 2576 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 3750656
  • Decoded allocation: 100000000 bytes
  • File bytes per image: 7000000 (Pumpkin cap; 7,000,000 file bytes encode to at most 9,333,336 base64 bytes, below the documented 10 MB)
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 28 px patches, budget 4784 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Long-running agentic coding and knowledge work. Newer and cheaper than Opus 5, still below Fable.

Standard Messages output limit; thinking shares output budget.

Adaptive thinking is always on; forced tool choice is unsupported. Default effort is medium.

Cache read is 0.05x base input, not the usual 0.1x.

Verified Sep 22, 2026

https://platform.claude.com/docs/en/models/opus-5-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing
Claude Opus 5anthropic · defaultclaude-opus-5Editorial skill: 8 / 10

Context: 1000000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 600

Formats and limits

Provider formats: image/jpeg, image/png, image/gif, image/webp

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: 8000 / 8000

Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified.

Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images.

Long edge: 2576 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 3750656
  • Decoded allocation: 100000000 bytes
  • File bytes per image: 7000000 (Pumpkin cap; 7,000,000 file bytes encode to at most 9,333,336 base64 bytes, below the documented 10 MB)
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 28 px patches, budget 4784 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Strong implementation and review when Fable is overkill: scoped features, correctness-sensitive edits, production-ready diffs.

Standard Messages output limit; thinking shares output budget.

Verified Sep 22, 2026

https://platform.claude.com/docs/en/models/opus-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing
Claude Sonnet 5anthropic · defaultclaude-sonnet-5Editorial skill: 6 / 10

Context: 1000000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 600

Formats and limits

Provider formats: image/jpeg, image/png, image/gif, image/webp

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: 8000 / 8000

Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified.

Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images.

Long edge: 2576 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 3750656
  • Decoded allocation: 100000000 bytes
  • File bytes per image: 7000000 (Pumpkin cap; 7,000,000 file bytes encode to at most 9,333,336 base64 bytes, below the documented 10 MB)
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 28 px patches, budget 4784 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Everyday coding: routine features, refactors with a known shape, tests, docs. Mid-tier default for Anthropic.

Standard Messages output limit; thinking shares output budget.

Verified Sep 22, 2026

https://platform.claude.com/docs/en/models/sonnet-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing
Claude Haiku 4.5anthropic · defaultclaude-haiku-4-5Editorial skill: 4 / 10

Context: 200000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 64000

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 100

Formats and limits

Provider formats: image/jpeg, image/png, image/gif, image/webp

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: 8000 / 8000

Direct Claude API: 10 MB base64-encoded per image; 32 MB whole request. MB byte definition unspecified.

Animation uses the first frame. More than 20 images has stricter dimensions; Pumpkin caps requests at 20 images.

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 1229312
  • Decoded allocation: 100000000 bytes
  • File bytes per image: 7000000 (Pumpkin cap; 7,000,000 file bytes encode to at most 9,333,336 base64 bytes, below the documented 10 MB)
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 28 px patches, budget 1568 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Cheap, short jobs: mechanical edits, triage, titles, compaction, apply this exact change. Not for design or long hunts.

Standard Messages output limit; thinking shares output budget.

Verified Sep 22, 2026

https://platform.claude.com/docs/en/models/haiku-4-5/overviewhttps://platform.claude.com/docs/en/build-with-claude/visionhttps://platform.claude.com/docs/en/about-claude/pricing
Grok 4.7xai · defaultgrok-4.7Editorial skill: 7 / 10

Context: 500000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: low, medium, high, xhigh

Supported: Yes

Count: Unknown

Formats and limits

Provider formats: image/jpeg, image/png

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

20 MiB per image; decoded versus encoded accounting unspecified. No image-count limit stated, but context limits still apply.

Hard pixel dimensions and numeric provider resize recommendation unknown.

Output request cap: 128000 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: xAI flagship for coding and agent work. Same price as 4.6, stronger on tools and harder implementation.

Provider documents no separate text output limit. maxOutput=128000 is Pumpkin request policy, not a provider maximum.

Reasoning cannot be disabled.

Verified Sep 22, 2026

https://docs.x.ai/developers/models/grok-4.7https://docs.x.ai/developers/model-capabilities/images/understandinghttps://docs.x.ai/developers/pricing
Grok 4.6xai · defaultgrok-4.6Editorial skill: 6 / 10

Context: 500000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: low, medium, high, xhigh

Supported: Yes

Count: Unknown

Formats and limits

Provider formats: image/jpeg, image/png

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

20 MiB per image; decoded versus encoded accounting unspecified. No image-count limit stated, but context limits still apply.

Hard pixel dimensions and numeric provider resize recommendation unknown.

Output request cap: 128000 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Fast workhorse: investigation, evidence collection, intermediate code, image analysis. Good value; less stubborn than Fable on long stuck turns.

Provider documents no separate text output limit. maxOutput=128000 is Pumpkin request policy, not a provider maximum.

Reasoning cannot be disabled.

Verified Sep 22, 2026

https://docs.x.ai/developers/models/grok-4.6https://docs.x.ai/developers/model-capabilities/images/understandinghttps://docs.x.ai/developers/pricing
GPT-5.6 Solopenai · defaultgpt-5.6-solEditorial skill: 5 / 10

Context: 1050000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Yes

Effort values: none, low, medium, high, xhigh, max

Supported: Yes

Count: 1500

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified.

Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity.

Long edge: 2048 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2560000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 32 px patches, budget 2500 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Solid mid OpenAI: focused features, bug fixes with a known cause, normal coding loops.

Maximum input 922000 tokens; reasoning and visible output share the output budget.

Promotional prices at least through 2026-11-21; subsequent rates unknown.

Verified Sep 22, 2026

https://developers.openai.com/api/docs/models/gpt-5.6-solhttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing
GPT-5.6 Terraopenai · defaultgpt-5.6-terraEditorial skill: 4 / 10

Context: 1050000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Yes

Effort values: none, low, medium, high, xhigh, max

Supported: Yes

Count: 1500

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified.

Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity.

Long edge: 2048 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2560000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 32 px patches, budget 2500 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Small, well-specified work: one-file edits, straightforward tests, low-stakes helpers.

Maximum input 922000 tokens; reasoning and visible output share the output budget.

Verified Sep 22, 2026

https://developers.openai.com/api/docs/models/gpt-5.6-terrahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing
GPT-5.6 Lunaopenai · defaultgpt-5.6-lunaEditorial skill: 2 / 10

Context: 1050000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Yes

Effort values: none, low, medium, high, xhigh, max

Supported: Yes

Count: 1500

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified.

Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity.

Long edge: 2048 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2560000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 32 px patches, budget 2500 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Cheapest OpenAI: tiny mechanical tasks, classification, short rewrites. Skip anything that needs judgment.

Maximum input 922000 tokens; reasoning and visible output share the output budget.

Verified Sep 22, 2026

https://developers.openai.com/api/docs/models/gpt-5.6-lunahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing
GPT-6 Astraopenai · defaultgpt-6-astraEditorial skill: 10 / 10

Context: 1050000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Yes

Effort values: low, medium, high, xhigh, max

Supported: Yes

Count: 1500

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

GIF must be nonanimated. Total request payload up to 512 MB; byte accounting unspecified.

Original/auto preserves resolution subject to a 65535px preprocessing cap; more than 30000 processed 32px patches per image is rejected. Pumpkin chooses a smaller high-detail patch budget, not original-detail fidelity.

Long edge: 65535 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2560000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 32 px patches, budget 2500 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: OpenAI flagship: hard implementation and long context, same class as Fable. Use when the problem is large or ambiguous.

Maximum input 922000 tokens; reasoning and visible output share the output budget.

Tool calling requires Responses; Chat Completions supports text only.

Verified Sep 22, 2026

https://developers.openai.com/api/docs/models/gpt-6-astrahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing
Muse Spark 1.3meta · defaultmuse-spark-1.3Editorial skill: 6 / 10

Context: 1048576 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: Unknown separate numeric maximum

Pumpkin output request cap: 65536 tokens, not a documented provider maximum.

Streaming: Yes

Effort values: minimal, low, medium, high, xhigh, max

Supported: Yes

Count: 20

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Images as base64 data URLs or Files API references; audio understanding is degraded on 1.3 (use 1.2), and GIF uses the first frame.

Meta documents no per-image or per-request image byte limits; Pumpkin applies its own conservative client policy rather than trusting an unspecified vendor cap.

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png
  • 32 px patches, budget 2500 per image

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Meta's agentic reasoning model: multi-step tool loops and coding, 1M context, cheap cached input.

Reasoning cannot be turned off (effort "none" returns HTTP 400); reasoning shares the output budget with visible text.

Meta documents no per-request output ceiling; 65536 is a safe request cap below the 1,048,576 context window (input and output share one budget).

Verified Sep 22, 2026

https://dev.meta.ai/docs/modelshttps://dev.meta.ai/docs/pricing-rate-limitshttps://dev.meta.ai/docs/reasoninghttps://dev.meta.ai/docs/image-understanding
DeepSeek V4 Pro 0813fireworks · opt-inaccounts/fireworks/models/deepseek-v4-pro-0813Editorial skill: 6 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: Strongest DeepSeek: hard coding, agents and long reasoning; Fireworks' top open pick.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://app.fireworks.ai/models/fireworks/deepseek-v4-pro-0813https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Kimi K3fireworks · opt-inaccounts/fireworks/models/kimi-k3Editorial skill: 6 / 10Availability: available

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Kimi K3 flagship: hard coding and research agents, image input, 1M context; the priciest here.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/kimi-k3https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
GLM 5.3fireworks · opt-inaccounts/fireworks/models/glm-5p3Editorial skill: 6 / 10Availability: available

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: Newest GLM flagship: coding, reasoning and agent tool use; 1M context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/glm-5p3https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
GLM 5.2fireworks · opt-inaccounts/fireworks/models/glm-5p2Editorial skill: 5 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: GLM 5.2 flagship: coding and agent tool use; 1M context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/glm-5p2https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Kimi K2.7 Codefireworks · opt-inaccounts/fireworks/models/kimi-k2p7-codeEditorial skill: 6 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 262144 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Kimi K2.7 Code: coding-tuned agent model with image input; 262K context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/kimi-k2p7-codehttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Kimi K2.6fireworks · opt-inaccounts/fireworks/models/kimi-k2p6Editorial skill: 5 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 262144 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Kimi K2.6: agents with tool use, images and documents; 262K context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/kimi-k2p6https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Qwen 3.8 2.4T A95Bfireworks · opt-inaccounts/fireworks/models/qwen3p8-2p4t-a95bEditorial skill: 5 / 10Availability: unavailable

Context: 262144 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: Qwen 3.8 (2.4T MoE, 95B active): strong general coding and reasoning; 262K context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Public model page says serverless not supported. No substitution of qwen3p8-max pricing.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/qwen3p8-2p4t-a95bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
MiniMax M3fireworks · opt-inaccounts/fireworks/models/minimax-m3Editorial skill: 5 / 10Availability: available

Context: 512000 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: MiniMax M3: fast coding and agent model, good value; 512K context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Exact serving-model feature metadata says no images despite generic multimodal architecture prose.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/minimax-m3https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
DeepSeek V4.1 Flashfireworks · opt-inaccounts/fireworks/models/deepseek-v4p1-flashEditorial skill: 4 / 10Availability: available

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: DeepSeek V4.1 Flash with image input: quick coding and agent steps at low cost; 1M context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://app.fireworks.ai/models/fireworks/deepseek-v4p1-flashhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Nemotron 3 Ultra NVFP4fireworks · opt-inaccounts/fireworks/models/nemotron-3-ultra-nvfp4Editorial skill: 4 / 10Availability: available

Context: 262144 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: NVIDIA Nemotron 3 Ultra (preview): large reasoning model; 262K context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/nemotron-3-ultra-nvfp4https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
DeepSeek V4 Flash 0731fireworks · opt-inaccounts/fireworks/models/deepseek-v4-flash-0731Editorial skill: 3 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: Fast, cheap DeepSeek: extraction, summaries and everyday agent steps; 1M context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://app.fireworks.ai/models/fireworks/deepseek-v4-flash-0731https://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
DeepSeek V4 Flash Vision Experimentalfireworks · opt-inaccounts/fireworks/models/deepseek-v4-flash-vision-expEditorial skill: 3 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: DeepSeek V4 Flash with experimental image input; cheap, 1M context.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://app.fireworks.ai/models/fireworks/deepseek-v4-flash-vision-exphttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
GPT OSS 120Bfireworks · opt-inaccounts/fireworks/models/gpt-oss-120bEditorial skill: 3 / 10Availability: available

Context: 131072 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: OpenAI open-weight 120B: solid reasoning and tool use at low cost; no images.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/gpt-oss-120bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
GLM 5.3 Flashfireworks · opt-inaccounts/fireworks/models/glm-5p3-flashEditorial skill: 3 / 10Availability: available

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Small, cheap GLM with image input: quick edits and lookups.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/glm-5p3-flashhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Muse Glimmer 30Bfireworks · opt-inaccounts/fireworks/models/muse-glimmer-30bEditorial skill: 3 / 10Availability: availableServerless retirement: Sep 25, 2026

Context: 131072 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Small 30B always-on agent model with image input: cheap and quick for simple tasks.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Serverless retirement announced for 2026-09-25; exact cutoff time unknown. Dedicated deployments are excluded.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/muse-glimmer-30bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Nemotron Lightning 3.5 30B A3Bfireworks · opt-inaccounts/fireworks/models/nemotron-lightning-3p5-30b-a3bEditorial skill: 2 / 10Availability: available

Context: 262144 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: No

Discovery reports image input unsupported; curated support is conservatively disabled.

Output request cap: 16384 tokens (also bounded by context).

Image policy: Not applicable / Unknown.

Details and sources

Editorial guidance: Cheapest fast agent model (30B, 3B active): small edits, descriptions, triage.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/nemotron-lightning-3p5-30b-a3bhttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
Inklingfireworks · opt-inaccounts/fireworks/models/inklingEditorial skill: UnknownAvailability: available

Context: 1048576 tokens

Tools: Yes

Reasoning: Unknown

Output and reasoning details

Provider output: Unknown separate numeric maximum

Streaming: Unknown

Effort values: Unknown

Supported: Yes

Count: 30

Formats and limits

Provider formats: image/png, image/jpeg, image/gif, image/bmp, image/tiff, image/x-portable-pixmap

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

Aggregate base64 image data must be less than 10 MB; MB definition unspecified. URL images less than 5 MB and download within 1.5 seconds.

No numeric universal provider resize recommendation. Pumpkin uses its conservative resize policy.

Output request cap: 16384 tokens (also bounded by context).

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 9000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Fireworks serverless model with image input and a 1M context; capability not rated.

Provider output limit Unknown. thinkingMode=none and the 16384 output cap describe Pumpkin adapter policy, not absence of provider reasoning.

Exact context Unknown until metadata discovery; rounded public context labels are not treated as exact limits.

Verified Sep 22, 2026

https://fireworks.ai/models/fireworks/inklinghttps://docs.fireworks.ai/guides/querying-vision-language-modelshttps://docs.fireworks.ai/serverless/pricinghttps://docs.fireworks.ai/updates/changeloghttps://api.fireworks.ai/v1/accounts/fireworks/models
GPT-6 Solopenai · defaultgpt-6-solEditorial skill: 6 / 10

Context: 1050000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Yes

Effort values: none, low, medium, high, xhigh, max

Supported: Yes

Count: 1500

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

GIF must be nonanimated. The image guide allows up to 512 MB total request payload; exact byte accounting is unspecified. Model-specific image dimensions remain Unknown.

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Complex coding and agent work when Astra is more than you need. Cheaper than GPT-5.6 Sol.

Use Responses for built-in tools and function calling. Chat Completions function calling requires reasoning_effort=none.

Exact model-specific image resize limits are not stated in the reviewed image guide; Pumpkin uses its conservative client policy, not an inferred provider maximum.

Verified Sep 22, 2026

https://developers.openai.com/api/docs/models/gpt-6-solhttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing
GPT-6 Lunaopenai · defaultgpt-6-lunaEditorial skill: 3 / 10

Context: 1050000 tokens

Tools: Yes

Reasoning: Yes

Output and reasoning details

Provider output: 128000

Streaming: Yes

Effort values: none, low, medium, high, xhigh, max

Supported: Yes

Count: 1500

Formats and limits

Provider formats: image/png, image/jpeg, image/webp, image/gif

Bytes per image: Unknown

Aggregate encoded bytes: Unknown

Hard width / height: Unknown / Unknown

GIF must be nonanimated. The image guide allows up to 512 MB total request payload; exact byte accounting is unspecified. Model-specific image dimensions remain Unknown.

Long edge: 1568 px
Images: 20

Pumpkin image limits

Pumpkin-selected image limits, not provider maxima:

  • Pixel area: 2500000
  • Decoded allocation: 100000000 bytes
  • Aggregate encoded image data: 20000000 bytes
  • Output format: image/png

Preserve aspect ratio; never upscale. Static GIF/WebP can be normalized; animated GIF is rejected. Provider-listed BMP/TIFF/PPM does not imply Pumpkin decoder support. Unsupported inputs fail explicitly.

Details and sources

Editorial guidance: Cheapest OpenAI: high-volume tasks, classification, short rewrites. Not for design or long hunts.

Use Responses for built-in tools and function calling. Chat Completions function calling requires reasoning_effort=none.

Exact model-specific image resize limits are not stated in the reviewed image guide; Pumpkin uses its conservative client policy, not an inferred provider maximum.

Verified Sep 22, 2026

https://developers.openai.com/api/docs/models/gpt-6-lunahttps://developers.openai.com/api/docs/guides/images-visionhttps://developers.openai.com/api/docs/pricing