- Cap
min_ctxrouting at the observed token quota when it is smaller than the advertised context window, including Groq preview models. - Keep response quotas and rate-limit status separate for each model so one model cannot exclude or admit another model from the same provider.
- Preserve per-model quota data when refreshing shared OpenRouter credits.
- Document the quota cap and add regression coverage. All 229 unit tests pass.