feat: bringing provider governance to budget & limits - #3938
Conversation
|
Caution Review failedThe pull request is closed. ℹ️ Recent review info⚙️ Run configurationConfiguration used: Path: .coderabbit.yaml Review profile: ASSERTIVE Plan: Pro Run ID: 📒 Files selected for processing (13)
📝 WalkthroughWalkthroughProvider-level governance (budget and rate-limit configuration) is migrated from provider row foreign keys to wildcard model configs, enabling provider-scoped "all models" governance policies. A database migration moves existing provider governance to wildcard entries, in-memory policy evaluation is refactored to support tiered config collection, HTTP endpoints transition to wildcard config persistence, and UI/cache invalidation ensure consistency. ChangesProvider Governance Wildcard Model Config Migration
Sequence DiagramsequenceDiagram
participant Client
participant Handler as HTTP Handler
participant Storage as RDB ConfigStore
participant InMem as In-Memory Governance
Client->>Handler: POST /provider-governance (update)
Handler->>Storage: Validate provider exists
Handler->>Storage: Fetch or create provider wildcard model config
Handler->>Storage: Upsert/delete wildcard model config + budget/rate-limit lifecycle
Handler->>InMem: remove or reload model config (sync)
Handler-->>Client: updated provider governance response
Note over Client,InMem: Subsequent checks evaluate wildcard tier
Client->>InMem: CheckModelBudget (model, provider)
InMem->>InMem: collectModelConfigsFor (exact, normalized, provider-wildcard, global-wildcard)
InMem-->>Client: budget/rate-limit decision
Estimated code review effort🎯 4 (Complex) | ⏱️ ~60 minutes Possibly related PRs
Suggested reviewers
🚥 Pre-merge checks | ✅ 4 | ❌ 1❌ Failed checks (1 warning)
✅ Passed checks (4 passed)
✏️ Tip: You can configure your own custom pre-merge checks in the settings. ✨ Finishing Touches📝 Generate docstrings
🧪 Generate unit tests (beta)
Warning There were issues while running some tools. Please review the errors and either fix the tool's configuration or disable the tool if it's a critical failure. 🔧 golangci-lint (2.12.2)level=error msg="[linters_context] typechecking error: pattern ./...: directory prefix . does not contain main module or its selected dependencies" Comment |
This stack of pull requests is managed by Graphite. Learn more about stacking. |
Confidence Score: 1/5Not safe to merge — the governance package does not compile due to two call sites referencing a removed method, and the The governance plugin calls
Important Files Changed
Reviews (4): Last reviewed commit: "feat: bringing provider governance to bu..." | Re-trigger Greptile |
|
@coderabbitai full-review |
|
✅ Actions performedFull review triggered. |
There was a problem hiding this comment.
Actionable comments posted: 8
🤖 Prompt for all review comments with AI agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.
Inline comments:
In `@framework/configstore/migrations_test.go`:
- Around line 2441-2479: Add additional unit tests alongside
TestMigrationMigrateProviderGovernanceToModelConfigs to cover edge cases: create
a provider with only BudgetID (nil RateLimitID), a provider with only
RateLimitID (nil BudgetID), a provider with neither (ensure no wildcard config
created), and multiple providers (each gets its own wildcard config). For each
new test, call migrationMigrateProviderGovernanceToModelConfigs(ctx, db) and
assert expected TableModelConfig rows (query TableModelConfig by
scope/model_name/provider) and that provider TableProvider budget_id and
rate_limit_id are cleared as appropriate; also verify idempotency by re-running
the migration and checking no duplicates are created. Ensure tests reuse the
same setup pattern (setupTestDB, AutoMigrate for
tables.TableProvider/tables.TableModelConfig/tables.TableBudget/tables.TableRateLimit)
and unique IDs for budgets/rate-limits to avoid cross-test collisions.
In `@framework/configstore/migrations.go`:
- Around line 3983-4005: The rollback currently mutates and deletes canonical
provider-governance rows (see the tx query selecting tables.TableModelConfig
rows into wildcards and the subsequent updates to tables.TableProvider and
tx.Delete of mc.ID), which can overwrite live data; instead mark this migration
as non-rollbackable by removing these destructive Down steps and returning an
explicit irreversible error (e.g. return fmt.Errorf("irreversible migration:
provider-governance canonicalization cannot be rolled back")) from the
migration's Down/rollback handler so callers know the migration cannot be safely
reverted.
In `@framework/configstore/rdb.go`:
- Around line 1159-1176: The current deletion first snapshots
providerModelConfigs and then runs a second WHERE "provider = ?" delete, which
allows a race with concurrent CreateModelConfig/UpdateModelConfig; to fix,
delete exactly the snapped rows (use the collected providerModelConfigs' IDs in
a single DELETE WHERE id IN (...) so you only remove the rows you read) and
additionally serialize concurrent model-config writers by taking a row-level
lock on the provider in CreateModelConfig and UpdateModelConfig (use a FOR
UPDATE/Locking clause on the provider select before creating/updating model
configs) so no new provider-scoped rows can be inserted between your snapshot
and delete; reference providerModelConfigs, CreateModelConfig,
UpdateModelConfig, and the txDB delete call to locate the changes.
In `@framework/configstore/store.go`:
- Line 259: Clarify the method comment for GetProviderGovernanceModelConfigs to
state it returns only provider-specific wildcard model configs (i.e., entries
where provider != nil) and does not include global wildcard configs (provider ==
nil); update the docstring to explicitly say this so implementers know the
function should filter and return only provider-scoped "all models on a
provider" configurations (matching the downstream handler's expectations).
In `@plugins/governance/store.go`:
- Around line 1087-1126: GetBudgetAndRateLimitStatus and
CollectApplicableGovernanceIDs currently only check exact-model and model-only
rows so they miss provider-level wildcard configs ("*", provider) and global
wildcards ("*", nil); update both functions to call collectModelConfigsFor(ctx,
scope, scopeID, model, provider) and iterate the returned
[]*configstoreTables.TableModelConfig to build budgets, rate-limits and
governance ID lists (honoring the deduping/order already applied by
collectModelConfigsFor), handling provider==nil and preserving existing behavior
for model-only and exact-model entries. Ensure you replace any direct lookups
like findScopedModelOnlyConfig or gs.modelConfigs.Load for model entries inside
those methods with the collectModelConfigsFor call and use the returned slice to
compute status and attribution.
- Around line 1281-1299: The code currently overwrites
entityWiseRateLimits[modelConfigEntityKey(mc)] with a single-element slice,
losing multiple matching rate limits and causing CheckRateLimit to see only one
violation; change the logic in CheckModelRateLimit (and the similar block at
1327-1351) to append rate limits instead of replacing them: for each mc with
mc.RateLimitID, load the rateLimit and if non-nil do existing :=
entityWiseRateLimits[modelConfigEntityKey(mc)];
entityWiseRateLimits[modelConfigEntityKey(mc)] = append(existing, rateLimit) so
CheckRateLimit receives all matching limits and can correctly aggregate
violations into DecisionRateLimited when both tokens and requests are exceeded.
In `@transports/bifrost-http/handlers/governance.go`:
- Around line 3139-3144: Replace the inline removal check that examines only
req.RateLimit.TokenMaxLimit and req.RateLimit.RequestMaxLimit with a call to the
existing helper isRateLimitRemovalRequest to centralize semantics; update the
block that mutates mc.RateLimitID and mc.RateLimit so it runs when
isRateLimitRemovalRequest(req.RateLimit) returns true instead of the current
two-field check (ensure you keep the assignments rateLimitIDToDelete =
*mc.RateLimitID, mc.RateLimitID = nil, mc.RateLimit = nil), and verify this
behavioral change (reset-duration-only requests no longer count as removals) is
acceptable.
In `@ui/lib/store/apis/governanceApi.ts`:
- Line 550: Add a short explanatory comment above the invalidatesTags:
["ProviderGovernance"] occurrence in governanceApi.ts that mirrors the rationale
used at the other two spots (lines where ProviderGovernance is invalidated) —
mention the cross-entity relationship and why invalidating ProviderGovernance is
necessary for this endpoint; keep the phrasing and placement consistent with the
comments around the other invalidatesTags instances.
🪄 Autofix (Beta)
Fix all unresolved CodeRabbit comments on this PR:
- Push a commit to this branch (recommended)
- Create a new PR with the fixes
ℹ️ Review info
⚙️ Run configuration
Configuration used: Path: .coderabbit.yaml
Review profile: ASSERTIVE
Plan: Pro
Run ID: 4c379fe0-1870-41ba-bb5e-9abd8e53c73c
📒 Files selected for processing (13)
framework/configstore/migrations.goframework/configstore/migrations_test.goframework/configstore/rdb.goframework/configstore/rdb_test.goframework/configstore/store.goframework/configstore/tables/modelconfig.goplugins/governance/modelprovidergovernance_test.goplugins/governance/store.gotransports/bifrost-http/handlers/governance.gotransports/bifrost-http/lib/config_test.goui/app/workspace/model-limits/views/modelLimitSheet.tsxui/app/workspace/model-limits/views/modelLimitsTable.tsxui/lib/store/apis/governanceApi.ts
eb27aa8 to
0b1b2f8
Compare
f3ef789 to
d0ce4b3
Compare
0b1b2f8 to
d7b2251
Compare
d0ce4b3 to
46704bf
Compare
Merge activity
|
46704bf to
ea9049f
Compare
## Summary
Provider-level governance (budget and rate limit) has been migrated from `config_providers.budget_id / rate_limit_id` into `governance_model_configs` as wildcard rows with `scope='global'`, `model_name='*'`, and `provider=<name>`. This makes `governance_model_configs` the single source of truth for all governance enforcement, eliminating the separate provider-governance enforcement path.
## Changes
- **Database migration** (`migrationMigrateProviderGovernanceToModelConfigs`): Folds existing provider-level budget/rate-limit FK references into new `(global, *, <provider>)` model config rows, reusing the same budget/rate-limit rows. Provider FKs are then nulled out. The migration is idempotent and includes a rollback path.
- **`ModelConfigAllModels = "*"` sentinel**: Introduced as a named constant to represent "all models" in a model config row. The `"*"` sentinel is excluded from catalog-based model name normalization.
- **`collectModelConfigsFor`**: New helper that resolves all applicable model configs for a request across four tiers — exact model+provider, exact model (all providers), all models on this provider (`*:provider`), and all models on all providers (`*:nil`) — deduped by config ID. All budget/rate-limit check and usage-tracking paths now use this helper instead of duplicating two-tier lookup logic.
- **`modelConfigEntityKey`**: New helper that builds a stable entity key for a model config, replacing ad-hoc `fmt.Sprintf` strings scattered across check and update functions.
- **`GetProviderGovernanceModelConfigs`**: New store method that queries the wildcard model config rows backing provider governance, with budget and rate-limit preloads.
- **`DeleteProvider` cleanup**: When a provider is deleted, its associated wildcard model configs (and their owned budget/rate-limit rows) are now cleaned up as part of the transaction.
- **Provider governance HTTP handlers**: `getProviderGovernance`, `updateProviderGovernance`, and `deleteProviderGovernance` are rewritten to operate on wildcard model config rows instead of provider FK columns. The GET endpoint gains a `from_memory` query parameter to serve data from the in-memory governance store.
- **UI**: The model limits table renders `"*"` as `"All Models"`. The model limit sheet exposes an `allowAllOption` on the model selector. RTK Query cache tags are cross-invalidated so that changes to model configs refresh the provider governance view and vice versa.
## Type of change
- [ ] Bug fix
- [x] Feature
- [x] Refactor
- [ ] Documentation
- [ ] Chore/CI
## Affected areas
- [x] Core (Go)
- [x] Transports (HTTP)
- [ ] Providers/Integrations
- [x] Plugins
- [x] UI (React)
- [ ] Docs
## How to test
```sh
# Core/Transports
go test ./framework/configstore/... ./plugins/governance/...
# UI
cd ui
pnpm i
pnpm build
```
1. Start with a deployment that has providers with existing `budget_id` or `rate_limit_id` values. After the migration runs, verify those providers have `budget_id = NULL` and `rate_limit_id = NULL`, and that corresponding `(global, *, <provider>)` rows exist in `governance_model_configs` pointing to the same budget/rate-limit IDs.
2. Re-run the migration and confirm no duplicate wildcard rows are created (idempotency).
3. Make a request through a provider that had governance configured and confirm the budget/rate-limit is still enforced.
4. Use `GET /api/governance/providers` and `GET /api/governance/providers?from_memory=true` and confirm both return the expected provider governance data.
5. Use `PUT` and `DELETE` on `/api/governance/providers/{name}` and confirm the wildcard model config rows are created, updated, or removed accordingly, and that the Model Limits UI reflects the change without a manual refresh.
6. Delete a provider and confirm its wildcard model config and owned budget/rate-limit rows are removed.
## Breaking changes
- [x] Yes
- [ ] No
Provider governance is no longer stored in `config_providers.budget_id / rate_limit_id`. Any code or tooling that reads governance directly from the providers table will no longer find it there. All governance enforcement and management must go through `governance_model_configs`. The HTTP API surface is unchanged; the migration handles existing data automatically.
## Security considerations
Budget and rate-limit rows are reused (not duplicated) during migration. The rollback path restores FK references to the provider table and removes the wildcard model config rows, leaving no orphaned governance rows in either direction.
## Checklist
- [x] I read `docs/contributing/README.md` and followed the guidelines
- [x] I added/updated tests where appropriate
- [ ] I updated documentation where needed
- [x] I verified builds succeed (Go and UI)
- [ ] I verified the CI pipeline passes locally if applicable
<!-- This is an auto-generated comment: release notes by coderabbit.ai -->
## Summary by CodeRabbit
* **New Features**
* Enabled governance policy configuration for "All Models" on specific providers, allowing budget and rate-limit controls at the provider level.
* **UI Improvements**
* Model limits table now displays "All Models" label for improved readability.
* Added "All Models" selection option in model name picker for creating provider-scoped policies.
* **Tests**
* Added comprehensive test coverage for provider-scoped all-models governance functionality.
<!-- end of auto-generated comment: release notes by coderabbit.ai -->
## Summary
Provider-level governance (budget and rate limit) has been migrated from `config_providers.budget_id / rate_limit_id` into `governance_model_configs` as wildcard rows with `scope='global'`, `model_name='*'`, and `provider=<name>`. This makes `governance_model_configs` the single source of truth for all governance enforcement, eliminating the separate provider-governance enforcement path.
## Changes
- **Database migration** (`migrationMigrateProviderGovernanceToModelConfigs`): Folds existing provider-level budget/rate-limit FK references into new `(global, *, <provider>)` model config rows, reusing the same budget/rate-limit rows. Provider FKs are then nulled out. The migration is idempotent and includes a rollback path.
- **`ModelConfigAllModels = "*"` sentinel**: Introduced as a named constant to represent "all models" in a model config row. The `"*"` sentinel is excluded from catalog-based model name normalization.
- **`collectModelConfigsFor`**: New helper that resolves all applicable model configs for a request across four tiers — exact model+provider, exact model (all providers), all models on this provider (`*:provider`), and all models on all providers (`*:nil`) — deduped by config ID. All budget/rate-limit check and usage-tracking paths now use this helper instead of duplicating two-tier lookup logic.
- **`modelConfigEntityKey`**: New helper that builds a stable entity key for a model config, replacing ad-hoc `fmt.Sprintf` strings scattered across check and update functions.
- **`GetProviderGovernanceModelConfigs`**: New store method that queries the wildcard model config rows backing provider governance, with budget and rate-limit preloads.
- **`DeleteProvider` cleanup**: When a provider is deleted, its associated wildcard model configs (and their owned budget/rate-limit rows) are now cleaned up as part of the transaction.
- **Provider governance HTTP handlers**: `getProviderGovernance`, `updateProviderGovernance`, and `deleteProviderGovernance` are rewritten to operate on wildcard model config rows instead of provider FK columns. The GET endpoint gains a `from_memory` query parameter to serve data from the in-memory governance store.
- **UI**: The model limits table renders `"*"` as `"All Models"`. The model limit sheet exposes an `allowAllOption` on the model selector. RTK Query cache tags are cross-invalidated so that changes to model configs refresh the provider governance view and vice versa.
## Type of change
- [ ] Bug fix
- [x] Feature
- [x] Refactor
- [ ] Documentation
- [ ] Chore/CI
## Affected areas
- [x] Core (Go)
- [x] Transports (HTTP)
- [ ] Providers/Integrations
- [x] Plugins
- [x] UI (React)
- [ ] Docs
## How to test
```sh
# Core/Transports
go test ./framework/configstore/... ./plugins/governance/...
# UI
cd ui
pnpm i
pnpm build
```
1. Start with a deployment that has providers with existing `budget_id` or `rate_limit_id` values. After the migration runs, verify those providers have `budget_id = NULL` and `rate_limit_id = NULL`, and that corresponding `(global, *, <provider>)` rows exist in `governance_model_configs` pointing to the same budget/rate-limit IDs.
2. Re-run the migration and confirm no duplicate wildcard rows are created (idempotency).
3. Make a request through a provider that had governance configured and confirm the budget/rate-limit is still enforced.
4. Use `GET /api/governance/providers` and `GET /api/governance/providers?from_memory=true` and confirm both return the expected provider governance data.
5. Use `PUT` and `DELETE` on `/api/governance/providers/{name}` and confirm the wildcard model config rows are created, updated, or removed accordingly, and that the Model Limits UI reflects the change without a manual refresh.
6. Delete a provider and confirm its wildcard model config and owned budget/rate-limit rows are removed.
## Breaking changes
- [x] Yes
- [ ] No
Provider governance is no longer stored in `config_providers.budget_id / rate_limit_id`. Any code or tooling that reads governance directly from the providers table will no longer find it there. All governance enforcement and management must go through `governance_model_configs`. The HTTP API surface is unchanged; the migration handles existing data automatically.
## Security considerations
Budget and rate-limit rows are reused (not duplicated) during migration. The rollback path restores FK references to the provider table and removes the wildcard model config rows, leaving no orphaned governance rows in either direction.
## Checklist
- [x] I read `docs/contributing/README.md` and followed the guidelines
- [x] I added/updated tests where appropriate
- [ ] I updated documentation where needed
- [x] I verified builds succeed (Go and UI)
- [ ] I verified the CI pipeline passes locally if applicable
<!-- This is an auto-generated comment: release notes by coderabbit.ai -->
## Summary by CodeRabbit
* **New Features**
* Enabled governance policy configuration for "All Models" on specific providers, allowing budget and rate-limit controls at the provider level.
* **UI Improvements**
* Model limits table now displays "All Models" label for improved readability.
* Added "All Models" selection option in model name picker for creating provider-scoped policies.
* **Tests**
* Added comprehensive test coverage for provider-scoped all-models governance functionality.
<!-- end of auto-generated comment: release notes by coderabbit.ai -->
## Summary This PR bumps the Go toolchain version from `1.26.3` to `1.26.4` across all modules and CI workflows, and cuts a new release (`core` v1.5.17, `framework` v1.3.17, `transports` v1.5.9, `plugins/compat` v0.1.16, `plugins/governance` v1.5.17, and associated plugin versions) incorporating a large batch of features and fixes accumulated since the previous release. ## Changes - **Go 1.26.4** — Updated `go-version` in all GitHub Actions workflows (`e2e-tests`, `helm-release`, `pr-tests`, `release-cli`, `release-pipeline`, `snyk`) and all `go.mod` files (core, framework, transports, cli, all plugins, examples, and test modules). - **Core (v1.5.17)** — OpenAI compaction support, multi-customer logs and usage tracking, multiple team/business unit support, `request_headers` wildcard pattern capture for OTel and Maxim plugins, xAI `x_search` tool, fetch URL validation with SSRF hardening, `file://` pricing URL scheme, virtual key provider fan-out filtering, and a broad set of fixes including Anthropic prompt cache key, empty thinking block stripping, OpenAI stream usage event cleanup, Gemini numeric schema constraints, stale connection retries, Azure Claude diagnostic strip, and passthrough budget handling. - **Framework (v1.3.17)** — Scope-aware budgets and limits wired from model configs, provider-level governance, multiple customer budget support with `calendar_aligned` windows, paginated virtual key fetch, `config.json` source-of-truth flow, FTS index cap reduction, sync worker drift fix, cascade deletes for model configs, and high-scale virtual key flow improvements. - **Transports (v1.5.9)** — Full changelog covering all of the above plus UI improvements (log navigation, customer detail sheet, `BudgetDisplay` component, inline loading shell, materialized view alias), SCIM provisioning fields, Helm/config schema additions (`roles`, `per_user_oauth`), client IP resolution from forwarded headers, and dependency upgrades (`recharts` to 3.8.1, `golang.org/x` CVE remediation). - **Plugins** — `governance` v1.5.17 adds team budget/rate-limit exporters, ghost node reconciliation fix, and VK double usage counting fix; `logging` v1.5.17 adds wildcard header capture and file attachment rendering; `otel` v1.2.17 adds `disable_content_logging` and multiple collectors support; `maxim` v1.6.17 adds `request_headers` wildcard capture; `compat` v0.1.16 fixes `max_tokens` preservation during param filtering. ## Type of change - [ ] Bug fix - [x] Feature - [ ] Refactor - [ ] Documentation - [x] Chore/CI ## Affected areas - [x] Core (Go) - [x] Transports (HTTP) - [x] Providers/Integrations - [x] Plugins - [x] UI (React) - [ ] Docs ## How to test ```sh # Verify Go version go version # should report go1.26.4 # Run core tests cd core && go test ./... # Run framework tests cd framework && go test ./... # Run transports tests cd transports && go test ./... # Run plugin tests cd plugins/governance && go test ./... cd plugins/logging && go test ./... cd plugins/otel && go test ./... # UI cd ui pnpm i pnpm build pnpm test ``` ## Breaking changes - [ ] Yes - [x] No ## Related issues #4053, #4066, #4041, #4012, #3976, #3947, #3991, #4045, #3957, #3938, #3937, #3939, #3981, #3998, #3997, #4092, #4091, #4079, #4080, #4086, #3929, #3994, #4028, #3970, #3919, #3861, #3664, #3999, #4088, #4070, #4051, #4043, #4057, #4023, #3941, #3955, #4024, #3956, #3967, #3925, #3992, #3900 ## Security considerations - Fetch URL validation hardened against SSRF by tightening IP checks for private networks and link-local addresses (#4092, #3947, #3991). - Transitive `golang.org/x` dependencies (crypto, net, sys, text) bumped to address Docker Scout CVEs (#3900). ## Checklist - [x] I read `docs/contributing/README.md` and followed the guidelines - [x] I added/updated tests where appropriate - [x] I updated documentation where needed - [x] I verified builds succeed (Go and UI) - [x] I verified the CI pipeline passes locally if applicable <!-- This is an auto-generated comment: release notes by coderabbit.ai --> ## Summary by CodeRabbit * **New Features** * OpenAI compaction, multi-customer/team logstore support, request-header wildcard capture, enhanced governance (provider-level & scope-aware limits), disable-content-logging option, support for multiple OpenTelemetry collectors, SSRF hardening and URL validation. * **Chores** * Bumped Go toolchain across modules and updated component/plugin version releases. <!-- end of auto-generated comment: release notes by coderabbit.ai -->
## Summary
Provider-level governance (budget and rate limit) has been migrated from `config_providers.budget_id / rate_limit_id` into `governance_model_configs` as wildcard rows with `scope='global'`, `model_name='*'`, and `provider=<name>`. This makes `governance_model_configs` the single source of truth for all governance enforcement, eliminating the separate provider-governance enforcement path.
## Changes
- **Database migration** (`migrationMigrateProviderGovernanceToModelConfigs`): Folds existing provider-level budget/rate-limit FK references into new `(global, *, <provider>)` model config rows, reusing the same budget/rate-limit rows. Provider FKs are then nulled out. The migration is idempotent and includes a rollback path.
- **`ModelConfigAllModels = "*"` sentinel**: Introduced as a named constant to represent "all models" in a model config row. The `"*"` sentinel is excluded from catalog-based model name normalization.
- **`collectModelConfigsFor`**: New helper that resolves all applicable model configs for a request across four tiers — exact model+provider, exact model (all providers), all models on this provider (`*:provider`), and all models on all providers (`*:nil`) — deduped by config ID. All budget/rate-limit check and usage-tracking paths now use this helper instead of duplicating two-tier lookup logic.
- **`modelConfigEntityKey`**: New helper that builds a stable entity key for a model config, replacing ad-hoc `fmt.Sprintf` strings scattered across check and update functions.
- **`GetProviderGovernanceModelConfigs`**: New store method that queries the wildcard model config rows backing provider governance, with budget and rate-limit preloads.
- **`DeleteProvider` cleanup**: When a provider is deleted, its associated wildcard model configs (and their owned budget/rate-limit rows) are now cleaned up as part of the transaction.
- **Provider governance HTTP handlers**: `getProviderGovernance`, `updateProviderGovernance`, and `deleteProviderGovernance` are rewritten to operate on wildcard model config rows instead of provider FK columns. The GET endpoint gains a `from_memory` query parameter to serve data from the in-memory governance store.
- **UI**: The model limits table renders `"*"` as `"All Models"`. The model limit sheet exposes an `allowAllOption` on the model selector. RTK Query cache tags are cross-invalidated so that changes to model configs refresh the provider governance view and vice versa.
## Type of change
- [ ] Bug fix
- [x] Feature
- [x] Refactor
- [ ] Documentation
- [ ] Chore/CI
## Affected areas
- [x] Core (Go)
- [x] Transports (HTTP)
- [ ] Providers/Integrations
- [x] Plugins
- [x] UI (React)
- [ ] Docs
## How to test
```sh
# Core/Transports
go test ./framework/configstore/... ./plugins/governance/...
# UI
cd ui
pnpm i
pnpm build
```
1. Start with a deployment that has providers with existing `budget_id` or `rate_limit_id` values. After the migration runs, verify those providers have `budget_id = NULL` and `rate_limit_id = NULL`, and that corresponding `(global, *, <provider>)` rows exist in `governance_model_configs` pointing to the same budget/rate-limit IDs.
2. Re-run the migration and confirm no duplicate wildcard rows are created (idempotency).
3. Make a request through a provider that had governance configured and confirm the budget/rate-limit is still enforced.
4. Use `GET /api/governance/providers` and `GET /api/governance/providers?from_memory=true` and confirm both return the expected provider governance data.
5. Use `PUT` and `DELETE` on `/api/governance/providers/{name}` and confirm the wildcard model config rows are created, updated, or removed accordingly, and that the Model Limits UI reflects the change without a manual refresh.
6. Delete a provider and confirm its wildcard model config and owned budget/rate-limit rows are removed.
## Breaking changes
- [x] Yes
- [ ] No
Provider governance is no longer stored in `config_providers.budget_id / rate_limit_id`. Any code or tooling that reads governance directly from the providers table will no longer find it there. All governance enforcement and management must go through `governance_model_configs`. The HTTP API surface is unchanged; the migration handles existing data automatically.
## Security considerations
Budget and rate-limit rows are reused (not duplicated) during migration. The rollback path restores FK references to the provider table and removes the wildcard model config rows, leaving no orphaned governance rows in either direction.
## Checklist
- [x] I read `docs/contributing/README.md` and followed the guidelines
- [x] I added/updated tests where appropriate
- [ ] I updated documentation where needed
- [x] I verified builds succeed (Go and UI)
- [ ] I verified the CI pipeline passes locally if applicable
<!-- This is an auto-generated comment: release notes by coderabbit.ai -->
## Summary by CodeRabbit
* **New Features**
* Enabled governance policy configuration for "All Models" on specific providers, allowing budget and rate-limit controls at the provider level.
* **UI Improvements**
* Model limits table now displays "All Models" label for improved readability.
* Added "All Models" selection option in model name picker for creating provider-scoped policies.
* **Tests**
* Added comprehensive test coverage for provider-scoped all-models governance functionality.
<!-- end of auto-generated comment: release notes by coderabbit.ai -->
## ✨ Features - **OpenAI Compaction** — Added OpenAI conversation compaction support across core, framework, logging, and the API surface (#4053) - **Multi-Customer & Org Hierarchy** — Logs and usage tracking now support multiple customers, teams, and business units, including business unit CRUD, team assignment, and governance endpoints in the OpenAPI spec (#4066, #4041, #4082) - **Provider-Level Governance** — Budgets & limits are now scope-aware and can be applied at the virtual-key top level and per provider, wired from the model configs table, with UI filters for scope and providers (#3938, #3937, #3939, #3981, #3962) - **Customer Budgets** — Customers support multiple budgets and `calendar_aligned` budget windows (#3998, #3997) - **Virtual Key Attribution & Controls** — Added a `created_by` user attribution column and a `blacklisted_models` column for virtual key provider configs (#3672, #3653) - **Request Header Capture** — OTel and Maxim observability plugins capture `request_headers` by pattern, with wildcard support (e.g. `x-custom-*`); logging gained the same wildcard header capture (#4012, #3958) - **OTel Content Controls & Collectors** — New `disable_content_logging` option drops message/tool content from exported spans, plus support for multiple OTel collectors (#4064, #3894) - **xAI x_search** — Added xAI `x_search` tool support (#3976) - **URL Validation** — Added fetch URL validation with private-network configuration and link-local blocking (#3947, #3991) - **File Scheme Pricing URLs** — Pricing source URLs now accept the `file://` scheme for air-gapped and self-hosted deployments (#4045) - **Paginated Virtual Keys** — Virtual key fetching is paginated to handle deployments with very large numbers of keys (#3957) - **Client IP Resolution** — Resolve client IP from `X-Forwarded-For`/`X-Real-IP` headers - **SCIM Provisioning** — Added `attributeType`/`attributeValue` SCIM provisioning fields - **Helm/Config Schema** — Added `roles` RBAC governance config and `per_user_oauth` MCP auth to the Helm chart and config schema (#4004, #4009) - **Log Navigation UI** — Added a "View logs" menu item to customer, team, and virtual key tables, clickable links in log detail views, a customer detail sheet, and a reusable `BudgetDisplay` component (#4073, #4054, #4026, #4055) - **Faster First Paint** — Added an inline loading shell to `#root` before React mounts (#4063) - **Materialized View Alias** — Added an `alias` column to the materialized view with filter support (#4078) ## 🐞 Fixed - **Fetch URL IP Checks** — Hardened fetch URL IP checks against SSRF (#4092) - **Mantle Model Matching** — Broadened Mantle model matching to all `gpt` variants (#4091) - **Empty Thinking Blocks** — Strip thinking blocks when the signature is empty (#4079) - **OpenAI Stream Usage** — Removed usage from the `responses.created` event in the OpenAI stream (#4080) - **Prompt Cache Key** — Set the prompt cache key from the Anthropic integration (#4086) - **Upstream Failure Status** — Map upstream connection failures to 502 instead of 400 (#3929) (thanks [@chris-colinsky](https://github.com/chris-colinsky)!) - **Gemini Schema Constraints** — Accept numeric schema integer constraints for Gemini (#3994) (thanks [@yanhao98](https://github.com/yanhao98)!) - **Files Provider Param** — Accept the `?provider=` query param on `GET /v1/files` (#3971) (thanks [@alexef](https://github.com/alexef)!) - **Optional Batch Model** — Made the `model` field optional on `POST /v1/batches` (#3973) (thanks [@alexef](https://github.com/alexef)!) - **Helm Azure Config** — Added missing `azure_key_config` fields to the Helm schema (#3996) (thanks [@axelray-dev](https://github.com/axelray-dev)!) - **Text Completion Chunk Model** — Added the missing `Model` field to `TextCompletionChunkResponse` (#3970) (thanks [@kuishou68](https://github.com/kuishou68)!) - **MCP Inline stdio Env** — MCP stdio server configs accept inline environment variable assignments (#3861) (thanks [@Shushmitaaaa](https://github.com/Shushmitaaaa)!) - **Orphaned Tool Results** — Orphaned tool results in the OpenAI to Anthropic conversion flow are no longer rejected by the Anthropic API (#3919) - **Node Usage Reconciliation** — Added a monotonic `inc_number` log cursor so node usage reconciliation does not skip late async log writes (#3664) - **Bedrock Output Assessments** — Corrected the type of `outputAssessments` in Bedrock responses (#4028) - **Model Pool Pricing Reloads** — Preserve non-pricing model pool entries across pricing reloads (#3999) - **Ghost Node Reconciliation** — Replicate the VK hierarchy flow for ghost node reconciliation (#4088) - **VK Double Usage Counting** — Fixed double usage counting when creating a virtual key (#4070) - **Model Config Lifecycle** — Cascade deletes for model configs and removal of stale in-memory model configs (#4051, #4043) - **FTS Index Cap** — Reduced the FTS index `left()` cap from 800k to 250k chars to stay within the tsvector limit (#4057) - **Sync Worker Drift** — Reduced the sync worker ticker period to 5m to prevent threshold drift (#4023) - **Passthrough** — Fixed passthrough budgets, gated passthrough models per VK, model extraction for Azure passthrough, and restricted fallbacks/provider selection to the VK boundary (#3941, #3988, #3983, #3924) - **Provider Response Headers** — Strip provider response headers and add a content-type filter (#3955, #4024) - **Stream Handling** — Drain non-SSE stream readers and retry stale connections (#3956, #3967) - **Azure Claude** — Strip Azure diagnostic property for Claude models (#3925) - **Compat max_tokens** — Preserve chat `max_tokens` during param filtering (#3992) - **Raw Request Flag** — Removed the raw request flag from providers that don't support it (#4058) - **UI Fixes** — Standardized page container layout, virtual key model configs UI, and dashboard chart tooltips (#4046, #4052, #4044) ## 🔧 Maintenance - **Dependency Upgrades** — Bumped transitive `golang.org/x` dependencies (crypto, net, sys, text) for Docker Scout CVE remediation and `recharts` to 3.8.1; cascaded version bumps across all modules (#3900, #4003)

Summary
Provider-level governance (budget and rate limit) has been migrated from
config_providers.budget_id / rate_limit_idintogovernance_model_configsas wildcard rows withscope='global',model_name='*', andprovider=<name>. This makesgovernance_model_configsthe single source of truth for all governance enforcement, eliminating the separate provider-governance enforcement path.Changes
migrationMigrateProviderGovernanceToModelConfigs): Folds existing provider-level budget/rate-limit FK references into new(global, *, <provider>)model config rows, reusing the same budget/rate-limit rows. Provider FKs are then nulled out. The migration is idempotent and includes a rollback path.ModelConfigAllModels = "*"sentinel: Introduced as a named constant to represent "all models" in a model config row. The"*"sentinel is excluded from catalog-based model name normalization.collectModelConfigsFor: New helper that resolves all applicable model configs for a request across four tiers — exact model+provider, exact model (all providers), all models on this provider (*:provider), and all models on all providers (*:nil) — deduped by config ID. All budget/rate-limit check and usage-tracking paths now use this helper instead of duplicating two-tier lookup logic.modelConfigEntityKey: New helper that builds a stable entity key for a model config, replacing ad-hocfmt.Sprintfstrings scattered across check and update functions.GetProviderGovernanceModelConfigs: New store method that queries the wildcard model config rows backing provider governance, with budget and rate-limit preloads.DeleteProvidercleanup: When a provider is deleted, its associated wildcard model configs (and their owned budget/rate-limit rows) are now cleaned up as part of the transaction.getProviderGovernance,updateProviderGovernance, anddeleteProviderGovernanceare rewritten to operate on wildcard model config rows instead of provider FK columns. The GET endpoint gains afrom_memoryquery parameter to serve data from the in-memory governance store."*"as"All Models". The model limit sheet exposes anallowAllOptionon the model selector. RTK Query cache tags are cross-invalidated so that changes to model configs refresh the provider governance view and vice versa.Type of change
Affected areas
How to test
budget_idorrate_limit_idvalues. After the migration runs, verify those providers havebudget_id = NULLandrate_limit_id = NULL, and that corresponding(global, *, <provider>)rows exist ingovernance_model_configspointing to the same budget/rate-limit IDs.GET /api/governance/providersandGET /api/governance/providers?from_memory=trueand confirm both return the expected provider governance data.PUTandDELETEon/api/governance/providers/{name}and confirm the wildcard model config rows are created, updated, or removed accordingly, and that the Model Limits UI reflects the change without a manual refresh.Breaking changes
Provider governance is no longer stored in
config_providers.budget_id / rate_limit_id. Any code or tooling that reads governance directly from the providers table will no longer find it there. All governance enforcement and management must go throughgovernance_model_configs. The HTTP API surface is unchanged; the migration handles existing data automatically.Security considerations
Budget and rate-limit rows are reused (not duplicated) during migration. The rollback path restores FK references to the provider table and removes the wildcard model config rows, leaving no orphaned governance rows in either direction.
Checklist
docs/contributing/README.mdand followed the guidelinesSummary by CodeRabbit
New Features
UI Improvements
Bug Fixes
Tests