Repository navigation
Devin plan-quota awareness, zero-write watchdog, and advise task-shape routing (0.12.3) - #33
Merged
Merged
Conversation
…k-shape routing (0.12.3) On a long Devin Pro session the shared daily plan quota can be exhausted, blocking every plan-included model (glm/swe/kimi) at once. Outsourcerer modeled only the paid ACU pool, so it read that exhaustion as a "mis-gate, not a real limit" and pointed the user back at the same dead lane. And `advise` recommended a pure-reasoning model for edit-run-verify work, with no glm-5.3-flash-high to prefer. Devin plan quota: - Detect the shared daily/weekly plan-quota exhaustion (distinct from a paid ACU refusal), mark the `dv` lane down until Devin's own stated reset, and steer OFF Devin (OpenRouter via --provider cc, or a native lane) instead of recommending another Devin plan model. - Parse Devin's "resets in 11h26m" / "resets at 09:00 UTC" wording into the real lane-down TTL. - Guard the dead `devin usage` call (the current CLI has no such subcommand); point every hint at the usage dashboard. - Per-session Devin-plan dispatch meter warns past OSRC_DEVIN_PLAN_JOBS_WARN that glm/swe/kimi/ deepseek share one daily bucket. Zero-write watchdog (lane-agnostic): - A live-but-writeless run (a model spending its budget on reasoning without touching a file) is surfaced as `no-progress-writes` past OSRC_NOPROGRESS_SECS, for mutating verbs only; read-only and text-only lanes are exempt. Surface-only unless OSRC_NOPROGRESS_KILL_SECS is set. advise: - glm-5.3 family incl. glm-5.3-flash-high added to the model table, bench map, and Devin resolver. - Edit-run-verify tasks classify as agentic; Kimi K3's near-frontier bump is reasoning-shaped only, so it no longer floats to #1 on edit/agentic work (a hard pure-reasoning task still picks it). - The recommendation prefers the cheaper capable variant within a family (glm-5.3-flash-high over glm-5.3-high) unless --effort max; OSRC_ADVISE_VARIANT_PREF=0 disables it. Scope: the plan-limit detection + failover is Devin-specific in this release; generalizing it to the other subscription harnesses (Warp, Droid, Cursor, Cline, native) is tracked as follow-up. Tests: three new suites (test_devin_plan_quota, test_nowrite_watchdog, test_advise_task_shape) plus an updated model-selection-parity suite; full conformance 141 passed, 0 failed.
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to subscribe to this conversation on GitHub.
Already have an account?
Sign in.
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
What
On a long Devin Pro session the shared daily plan quota can be exhausted, which blocks every plan-included model (glm/swe/kimi) at once. Outsourcerer modeled only the paid ACU pool, so it read that exhaustion as a "mis-gate, not a real limit" and pointed the user back at the same dead lane. Separately,
adviserecommended a pure-reasoning model for edit-run-verify work and had noglm-5.3-flash-highto prefer.Changes
Devin plan quota
dvlane down until Devin's own stated reset; steer off Devin (OpenRouter via--provider cc, or a native lane) instead of another Devin plan model.resets in 11h26m/resets at 09:00 UTCwording into the real lane-down TTL.devin usagecall (the current CLI has no such subcommand); hints point at the usage dashboard.OSRC_DEVIN_PLAN_JOBS_WARNthat glm/swe/kimi/deepseek share one daily bucket.Zero-write watchdog (lane-agnostic)
no-progress-writespastOSRC_NOPROGRESS_SECS, mutating verbs only; read-only and text-only lanes exempt. Surface-only unlessOSRC_NOPROGRESS_KILL_SECSis set.advise
glm-5.3-flash-highin the model table, bench map, and Devin resolver.--effort max;OSRC_ADVISE_VARIANT_PREF=0disables.Scope
Plan-limit detection + failover is Devin-specific in this release. Generalizing to the other subscription harnesses (Warp, Droid, Cursor, Cline, native) is tracked as follow-up. Proactive conservation (Claude/ChatGPT windows) and read-only transport-fallback are unchanged; the zero-write watchdog and
advisechanges are lane-agnostic.Tests
Three new suites (
test_devin_plan_quota,test_nowrite_watchdog,test_advise_task_shape) plus an updatedtest_model_selection_parity. Full conformance: 141 passed, 0 failed. Version parity green (0.12.3).