Skip to content

Reduce malformed managed ranking output within the existing token budget - #43

Merged
sarthakagrawal927 merged 1 commit into
mainfrom
agent/meme-lab-compact-ranking-20261005
Oct 5, 2026
Merged

sarthakagrawal927 merged 1 commit into
mainfrom
agent/meme-lab-compact-ranking-20261005

Conversation

@sarthakagrawal927

Copy link
Copy Markdown
Member

Live full-batch ranking intermittently returned invalid structured JSON or exhausted its 15-second deadline even after Gemini access-denial recovery. Request compact numeric score tuples instead of repeating field names for every candidate, and decode them into the unchanged normalized classifier response. Existing object answers remain accepted; invalid output still fails validation and uses the existing bounded retry/fallback.

Keep model: auto, all 30 candidates/labels, the 2,000-token ceiling, original deadline and request count. Add allowlisted finish-reason/token metadata for invalid JSON without logging prompts or provider content. The observed JSON failure is confirmed; provider-side truncation is not yet proven.

Validation: 15 focused tests and the full 220-test suite passed, including 30-candidate ordinal and perspective decoding, malformed tuples, truncation recovery and log privacy. Package validation, full public build and Wrangler deployment dry-run passed. Exact-source CI must pass before release; live strict qualification follows deployment. Refs #30 and sass-maker/free-ai#95.

@sarthakagrawal927
sarthakagrawal927 merged commit 16efc01 into main Oct 5, 2026
1 check passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant