Skip to content

release: v1.0.0-alpha.26 - #64

Merged
cativo23 merged 10 commits into
mainfrom
release/v1.0.0-alpha.26
Sep 30, 2026
Merged

cativo23 merged 10 commits into
mainfrom
release/v1.0.0-alpha.26

Conversation

@cativo23

Copy link
Copy Markdown
Owner

Fixed

  • OpenAI SDK auto-retry multiplying forensic-tier timeouts. See CHANGELOG.

Deploy notes

Sixth same-day-adjacent hotfix, following v1.0.0-alpha.20 through v1.0.0-alpha.25.

…migrations)

- App and worker container env vars all set (checked by name only)
- Public runtime config reached client (NUXT_PUBLIC_SUPABASE_URL in HTML)
- Storage buckets contracts and analysis-pdfs both exist, both private
- All 22 repo migrations recorded in production _migrations table, 0 pending
- Task 2: Carlos confirmed QA account provisioned, logged into connected
  Chrome profile; declined to create the credentials file, so AUTH_MODE=browser
  (the plan's own anticipated fallback) is used
- Task 3 step 1: public SUPABASE_URL/SUPABASE_ANON_KEY extracted from the
  production homepage (no login needed) and written to
  ~/.config/clarify-qa/phase13-ids (mode 600, values not echoed)
- Task 3 steps 2-8 blocked: AUTH_MODE=browser requires driving the existing
  authenticated Chrome tab via mcp__claude-in-chrome__* tools, which are not
  part of this gsd-executor subagent's tool set (confirmed via the
  claude-in-chrome skill this session) — no job submitted, no file uploaded,
  no credits spent; continuation needed from a context with browser tool access
…iling

Two consecutive Forensic-tier analyses failed in production with
"Request timed out" — but both took ~15 minutes to surface the error,
not the configured 10-minute timeout. The OpenAI SDK defaults to 2
automatic retries on failure (including timeout), and each retry
re-waits the full configured timeout, turning a single 10-minute
ceiling into an unbounded multiplier for a request that's already the
most expensive one this app makes (120k input / 30k output tokens,
gpt-6-astra).

Sets maxRetries: 0 so the configured timeout is a real wall-clock
ceiling, and raises the forensic timeout to 12 minutes (queue/worker
job timeouts raised to 13 minutes to stay above it) — a request this
large may legitimately need more headroom than 10 minutes even on a
single clean attempt, and retries provide little value for a
non-idempotent, expensive generation call anyway.

Verified locally: image builds successfully with the change.
fix(worker): disable OpenAI retry, raise forensic timeout
@cativo23
cativo23 merged commit be1f782 into main Sep 30, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant