Argon 1M output tokens vs Sume Format 120,000-character schema cap

A 1M-token output limit does not lift Sume's output_schema limits. The 120,000 characters bound the schema, and the fallback projection reads 8,000 characters.

5 min readSume
All posts

No: a model that can write 1 million output tokens does not let a Sume Format run return a bigger structured object. Google says Gemini 4 Argon has an output limit of 1 million tokens, up from 64,000 (read 2026-10-04). Sume's 120,000-character limit is not an output cap at all. It bounds the output_schema document you send, summed over every property name, key and string value, and the other schema limits are separate.

The two numbers measure different things, and mixing them up leads to schemas that are either too timid or rejected at submit.

What do Sume's limits apply to?

Sume's Structured output page lists four size limits. Each applies to the schema you bind to a run, and a violation is a 400 output_schema_invalid before anything runs, so nothing is charged.

Output schema size limits from Sume's Structured output docs, read 2026-10-04.
LimitValueViolation code
Nesting depth10 levelsmax_depth
Total properties5,000 across the whole documentmax_properties
Enum values1,000 per enummax_enum_values
Total string length120,000 characters over names, keys and string valuesmax_string_length

Where does a long model reply go in a Format run?

A Format run's result is mostly media: durable files on media.sume.com listed in artifacts[]. The orchestrator's closing text is part of output.text. When the run does not submit your object itself, Sume builds one afterwards from the generated media and the run's closing text, truncated to its first 8,000 characters. That fallback is the filled_by: "projection" path.

So a very long closing message is not a way to carry more data. If a field must be in output, require only what the Format actually makes, and keep long source material in input, which is written whole to a file in the run's workspace and is never truncated.

How should I design a schema for a long-output model?

Keep the schema small and the artifacts large.

  • Put media in SumeMediaFile# fields rather than inlining long text.
  • Make optional fields a type union with null, since an empty projection field reads as null.
  • Keep long description annotations short; they count against the 120,000-character budget.
  • Send the long brief in input; the first ~4,000 characters of instruction are carried as prompt text.

Does this change if the orchestrator changes?

The model field on a run selects the orchestrator only, per Create a run. The schema subset, the projection rule and the limits above are properties of the Format API, so they hold whichever model plans the run.

Sources

Related posts

More in Formats

All Formats posts

Written by Sume