coder

mirror of https://github.com/coder/coder.git synced 2026-06-04 05:28:20 +00:00

Author	SHA1	Message	Date
Danny Kopping	1dd0519a38	docs: clarify `max_connections` implications (#21596 ) Signed-off-by: Danny Kopping <danny@coder.com>	2026-01-21 22:10:12 +02:00
Danny Kopping	a14a22eb54	feat: support custom bedrock base url (#21582 ) Closes https://github.com/coder/aibridge/issues/126 Depends on https://github.com/coder/aibridge/pull/131 --------- Signed-off-by: Danny Kopping <danny@coder.com>	2026-01-21 12:48:56 +00:00
Ben Potter	6346eb7af8	docs: mention AI Governance add-on (#21592 ) Ironically, no AI was used to make this PR. --------- Co-authored-by: Matt Vollmer <matthewjvollmer@outlook.com>	2026-01-20 16:44:14 -06:00
Kacper Sawicki	ed679bb3da	feat(codersdk): add circuit breaker configuration support for aibridge (#21546 ) ## Summary Add circuit breaker support for AI Bridge to protect against cascading failures from upstream AI provider rate limits (HTTP 429, 503, and Anthropic's 529 overloaded responses). ## Changes - Add 5 new CLI options for circuit breaker configuration: - `--aibridge-circuit-breaker-enabled` (default: false) - `--aibridge-circuit-breaker-failure-threshold` (default: 5) - `--aibridge-circuit-breaker-interval` (default: 10s) - `--aibridge-circuit-breaker-timeout` (default: 30s) - `--aibridge-circuit-breaker-max-requests` (default: 3) - Update aibridge dependency to include circuit breaker support - Add tests for pool creation with circuit breaker providers ## Notes - Circuit breaker is disabled by default for backward compatibility - When enabled, applies to both OpenAI and Anthropic providers - Uses sony/gobreaker internally via the aibridge library ## Testing ``` make test RUN=TestPoolWithCircuitBreakerProviders ```	2026-01-20 14:59:29 +01:00
Rowan Smith	b163b4c950	feat: support bundle updates to enable pprof and telemetry collection (#21486 ) - Adds pprof collection support now that we have the listeners automatically starting (requires Coder server 2.28.0+, includes a version check). Collects heap, allocs, profile (30s), block, mutex, goroutine, threadcreate, trace (30s), cmdline, symbol. Performs capture for 30 seconds and emits a log line stating as such. Enable capture by supplying the `--pprof` flag or `CODER_SUPPORT_BUNDLE_PPROF` env var. Collection of pprof data from both coderd and the Coder agent occurs. - Adds collection of Prometheus metrics, also requires 2.28.0+ - Adds the ability to include a template in the bundle independently of supplying the details of a running workspace by supplying the `--template` flag or `CODER_SUPPORT_BUNDLE_TEMPLATE` env var - Captures a list of workspaces the user has access to. Defaults to a max of 10, configurable via `--workspaces-total-cap` / `CODER_SUPPORT_BUNDLE_WORKSPACES_TOTAL_CAP` - Collects additional stats from the coderd deployment (aggregated workspace/session metrics), as well as entitlements via license and dismissed health checks. created with help from mux	2026-01-20 10:28:52 +11:00
Susana Ferreira	a002fbbae6	refactor: avoid terminology collision with aibridge by renaming passthrough to tunneled (#21562 ) ## Description Renames "passthrough" to "tunneled" in aiproxy to avoid terminology collision with aibridge, which has its own passthrough concept. Follow-up from: https://github.com/coder/coder/pull/21512#discussion_r2698231778 --------- Co-authored-by: Danny Kopping <danny@coder.com>	2026-01-19 13:23:42 +00:00
Susana Ferreira	a406ed7cc5	feat: add upstream proxy support to aiproxy for passthrough requests (#21512 ) ## Description Adds upstream proxy support for AI Bridge Proxy passthrough requests. This allows aiproxy to forward non-allowlisted requests through an upstream proxy. Currently, the only supported configuration is when aiproxy is the first proxy in the chain (client → aiproxy → upstream proxy). ## Changes * Add `--aibridge-proxy-upstream` option to configure an upstream HTTP/HTTPS proxy URL for passthrough requests * Add `--aibridge-proxy-upstream-ca` option to trust custom CA certificates for HTTPS upstream proxies * Passthrough requests (non-allowlisted domains) are forwarded through the upstream proxy * MITM'd requests (allowlisted domains) continue to go directly to aibridge, not through the upstream proxy * Add tests for upstream proxy configuration and request routing Closes: https://github.com/coder/internal/issues/1204	2026-01-19 08:50:57 +00:00
Asher	4d414a0df7	feat: add --use-parameter-defaults flag (#21119 ) This is like `--yes`, but for parameter prompts.	2026-01-16 17:04:57 -09:00
Zach	ea465d4ea3	docs: add documentation for boundary audit logs (#21529 )	2026-01-16 13:04:06 -07:00
Yevhenii Shcherbina	fe68ec9095	chore: bump claude-code module version (#21527 ) - update boundary docs - bump claude-code module version - modify boundary policy for dogfood	2026-01-16 12:31:25 -05:00
Yevhenii Shcherbina	61961db41d	docs: update boundary docs (#21524 ) - update boundary docs - bump boundary version in dogfood	2026-01-15 15:07:40 -05:00
Ehab Younes	6683d807ac	refactor: add RFC-compliant enum types and use SDK as source of truth (#21468 ) Add comprehensive OAuth2 enum types to codersdk following RFC specifications: - OAuth2ProviderGrantType (RFC 6749) - OAuth2ProviderResponseType (RFC 6749) - OAuth2TokenEndpointAuthMethod (RFC 7591) - OAuth2PKCECodeChallengeMethod (RFC 7636) - OAuth2TokenType (RFC 6749, RFC 9449) - OAuth2RevocationTokenTypeHint (RFC 7009) - OAuth2ErrorCode (RFC 6749, RFC 7009, RFC 8707) Add OAuth2TokenRequest, OAuth2TokenResponse, OAuth2TokenRevocationRequest, and OAuth2Error structs to the SDK. Update OAuth2ClientRegistrationRequest, OAuth2ClientRegistrationResponse, OAuth2ClientConfiguration, and OAuth2AuthorizationServerMetadata to use typed enums instead of raw strings. This makes codersdk the single source of truth for OAuth2 types, eliminating duplication between SDK and server-side structs. Closes #21476	2026-01-15 12:41:28 +03:00
George K	0712faef4f	feat(enterprise): implement organization "disable workspace sharing" option (#21376 ) Adds a per-organization setting to disable workspace sharing. When enabled, all existing workspace ACLs in the organization are cleared and the workspace ACL mutation API endpoints return `403 Forbidden`. This complements the existing site-wide `--disable-workspace-sharing` flag by providing more granular control at the organization level. Closes https://github.com/coder/internal/issues/1073 (part 2) --------- Co-authored-by: Steven Masley <Emyrk@users.noreply.github.com>	2026-01-14 09:47:50 -08:00
Danny Kopping	7d5cd06f83	feat: add `aibridge` structured logging (#21492 ) Closes https://github.com/coder/internal/issues/1151 Sample: ``` [API] 2026-01-13 15:50:20.795 [info] coderd.aibridgedserver: interception started trace=8bb5a1d8eb10526cc46ad90f191bb468 span=a3e5b5da9546032a record_type=interception_start interception_id=97461880-4a6c-47c1-8292-3588dd715312 initiator_id=360c6167-a93a-4442-9c3e-f87a6d1cfb66 api_key_id=vg1sbUv97d provider=anthropic model=claude-opus-4-5-20251101 started_at="2026-01-13T15:50:20.790690781Z" metadata={} [API] 2026-01-13 15:50:23.741 [info] coderd.aibridgedserver: token usage recorded trace=8bb5a1d8eb10526cc46ad90f191bb468 span=a114f0cc3047296e record_type=token_usage interception_id=97461880-4a6c-47c1-8292-3588dd715312 msg_id=msg_01VJH1rYKspfun8BW29CrYEu input_tokens=10 output_tokens=8 created_at="2026-01-13T15:50:23.731587038Z" metadata={"cache_creation_input":53194,"cache_ephemeral_1h_input":0,"cache_ephemeral_5m_input":53194,"cache_read_input":0,"web_search_requests":0} [API] 2026-01-13 15:50:26.265 [info] coderd.aibridgedserver: token usage recorded trace=8bb5a1d8eb10526cc46ad90f191bb468 span=dbdafb563bff2c9c record_type=token_usage interception_id=97461880-4a6c-47c1-8292-3588dd715312 msg_id=msg_01VJH1rYKspfun8BW29CrYEu input_tokens=0 output_tokens=130 created_at="2026-01-13T15:50:26.254467904Z" metadata={} [API] 2026-01-13 15:50:26.268 [info] coderd.aibridgedserver: prompt usage recorded trace=8bb5a1d8eb10526cc46ad90f191bb468 span=da51887a757226fc record_type=prompt_usage interception_id=97461880-4a6c-47c1-8292-3588dd715312 msg_id=msg_01VJH1rYKspfun8BW29CrYEu prompt="list the jmia share price" created_at="2026-01-13T15:50:26.255299811Z" metadata={} [API] 2026-01-13 15:50:26.268 [info] coderd.aibridgedserver: interception ended trace=8bb5a1d8eb10526cc46ad90f191bb468 span=3fa25397705ee7c9 record_type=interception_end interception_id=97461880-4a6c-47c1-8292-3588dd715312 ended_at="2026-01-13T15:50:26.25555547Z" [API] 2026-01-13 15:50:26.269 [info] coderd.aibridgedserver: tool usage recorded trace=8bb5a1d8eb10526cc46ad90f191bb468 span=b54af90afc604d29 record_type=tool_usage interception_id=97461880-4a6c-47c1-8292-3588dd715312 msg_id=msg_01VJH1rYKspfun8BW29CrYEu tool=mcp__stonks__getStockPriceSnapshot input="{\"ticker\":\"JMIA\"}" server_url="" injected=false invocation_error="" created_at="2026-01-13T15:50:26.255164652Z" metadata={} ``` Structured logging is only enabled when `CODER_AIBRIDGE_STRUCTURED_LOGGING=true`. --------- Signed-off-by: Danny Kopping <danny@coder.com>	2026-01-14 17:26:08 +02:00
Sas Swart	ffa83a4ebc	docs: add documentation for coder script ordering (#21090 ) This Pull request adds documentation and guidance for the Coder script ordering feature. We: * explain the use case, benefits, and requirements. * provide example configuration snippets * discuss best practices and troubleshooting --------- Co-authored-by: Cian Johnston <cian@coder.com> Co-authored-by: DevCats <christofer@coder.com>	2026-01-14 14:40:38 +02:00
Andrew Aquino	0c5809726d	fix(docs): show dynamic parameters demo in local GIF instead of Imgur link (#21487 ) fixes this bug where the dynamic parameters demo GIF isn't viewable in the UK: <img width="720" height="798" alt="image" src="https://github.com/user-attachments/assets/757cd4fb-6b32-4db8-87fa-31a01588d69d" />	2026-01-13 09:31:32 -08:00
Susana Ferreira	74b6d12a8a	feat: implement selective MITM with configurable domain allowlist in aibridgeproxyd (#21473 ) ## Description Implements selective MITM (Man-in-the-Middle) in `aibridgeproxyd` so that only requests to allowlisted domains are intercepted and decrypted. Requests to all other domains are tunneled directly without decryption. ## Changes * New config option: `CODER_AIBRIDGE_PROXY_DOMAIN_ALLOWLIST` (default: `api.anthropic.com`,`api.openai.com`) * Selective MITM: Uses `goproxy.ReqHostIs()` to only intercept `CONNECT` requests to allowlisted hosts * Certificate caching: Now only generates/caches certificates for allowlisted domains * Validation: Startup fails if domain allowlist is empty or contains invalid entries Closes: https://github.com/coder/internal/issues/1182	2026-01-13 11:30:51 +00:00
Danny Kopping	49a42eff5c	feat: make database connection pool size configurable (#21403 ) Closes https://github.com/coder/coder/issues/21360 A few considerations/notes: - I've kept the number of conns to 10 in all other places, except coderd - which uses the config value - I opted to also make idle conns configurable; the greater the delta between max open and max idle, the more connection churn - Postgres maintains a [_process_ per connection](https://www.postgresql.org/docs/current/connect-estab.html), contrary to what the comment said previously - Operators should be able to tune this, since process churn can negatively affect OS scheduling - I've set the value to `"auto"` by default so it's not another knob one _has to_ twiddle, and sets max idle = max conns / 3 --------- Signed-off-by: Danny Kopping <danny@coder.com>	2026-01-13 10:50:57 +02:00
George K	cc2efe9e1f	feat(coderd/rbac): make organization-member a per-org system custom role (#21359 ) Migrated the built-in organization-member role to DB storage so it can be customized per org. Closes https://github.com/coder/internal/issues/1073 (part 1)	2026-01-12 18:19:19 -08:00
Kacper Sawicki	6ca70d3618	feat(cli): add --no-build flag to state push for state-only updates (#21374 ) ## Summary Adds a `--no-build` flag to `coder state push` that updates the Terraform state directly without triggering a workspace build. ## Use Case This enables state-only migrations, such as migrating Kubernetes resources from deprecated types (e.g., `kubernetes_config_map`) to versioned types (e.g., `kubernetes_config_map_v1`): ```bash coder state pull my-workspace > state.json terraform init terraform state rm -state=state.json kubernetes_config_map.example terraform import -state=state.json kubernetes_config_map_v1.example default/example coder state push --no-build my-workspace state.json ``` ## Changes - Add `PUT /api/v2/workspacebuilds/{id}/state` endpoint to update state without triggering a build - Add `UpdateWorkspaceBuildState` SDK method - Add `--no-build`/`-n` flag to `coder state push` - Add confirmation prompt (can be skipped with `--yes`/`-y`) since this is a potentially dangerous operation - Add test for `--no-build` functionality Fixes #21336	2026-01-12 15:16:59 +01:00
Yevhenii Shcherbina	1bfd776cb4	docs: add docs for boundary rules engine (#21471 ) Closes: https://github.com/coder/boundary/issues/146 - added docs for rules engine - move all boundary-related docs under new `boundary` directory	2026-01-09 15:04:51 -05:00
Jiachen Jiang	a09d85cc26	docs: provide guidance on shared workspaces (#21214 ) Co-authored-by: ケイラ <mckayla@hey.com>	2026-01-09 11:07:46 -08:00
Steven Masley	89f4d60e7b	chore: remove experiment "terraform-directory-reuse" (#21397 ) Experiment is no longer required, the new method will be released without an experiment and without a toggle Main PR is: https://github.com/coder/coder/pull/21398	2026-01-09 11:13:16 -06:00
Spike Curtis	4bc49ed6eb	docs: update scale architecture and add 10k user doc (#21454 ) Updates 2k, 3k docs to match previous changes to 1k ( #21362), including new database recommendations. Adds a 10k doc.	2026-01-09 08:16:11 +04:00
Yevhenii Shcherbina	1e8c292855	docs: update boundary docs (#21458 )	2026-01-08 15:11:03 -05:00
Cian Johnston	0f446f99dd	feat(cli): add logs cmd (#21430 ) This PR adds a command to view the provisioner and agent logs for a given workspace. Note: I did investigate using the existing `cliui` methods to tail the logs but they are tailored to a very specific use-case. Other changes: - Adds `Agents` to `dbfake.WorkspaceResponse` - Adds methods to generate provisioner and agent logs in `dbgen` --------- Co-authored-by: Steven Masley <Emyrk@users.noreply.github.com>	2026-01-08 09:58:10 +00:00
Atif Ali	989def7a94	docs: document coder_script resource (#21409 ) Co-authored-by: Claude Sonnet 4.5 <noreply@anthropic.com> Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>	2026-01-07 00:04:46 +05:00
Atif Ali	ef45ce4dfb	chore(docs): update example URL format in AI Bridge docs (#21435 ) We are using a mix of styles for the URL, so moving the example URLs to be code blocks. This is a cosmetic change. <table> <tr> <td> Before <td> After <tr> <td> <img width="712" height="607" alt="image" src="https://github.com/user-attachments/assets/05733e45-b83f-4801-9657-db72095a0300" /> <td> <img width="701" height="584" alt="image" src="https://github.com/user-attachments/assets/61418ea6-0882-4fb5-81fd-ec1a67473579" /> </table>	2026-01-06 21:14:13 +05:00
Yevhenii Shcherbina	2a7a33bb46	docs: update manifest in docs (#21434 ) Follow-up for https://github.com/coder/coder/pull/21420	2026-01-06 09:58:30 -05:00
Spike Curtis	ed6d41a5ef	docs: simplify 1k scale architecture and change db recommendation (#21362 ) DRAFT: I'd like feedback on this approach for 1k before I give the others the same treatment and add a 10k document. - Bumps database requirements to 8 vCPU, 30 GB memory. In our testing database was nearly always the bottleneck. (This could come back down again with improvements to how we use it.) - Removes specific machine type recommendations. - This only applies to VM-based deployments and many of our customers use Kubernetes. - The major clouds upgrade their machine teirs, so our recommendations go out of date - In its place we just give CPU and memory requirements - Removes API requests per second - It's not a metric that many operators will know until they are already operating - Our API requests vary wildly in cost depending on what they are - Replaces them with Users \| Running Workspaces \| Concurrent Builds - which represents our scale testing scenarios, and are easier for operators to reason about. - Removes specific advice about workspace sizing, instead gives the minimum specs for the agent - Gives Kubernetes resource request/limits in notes - Adds advice about not needing high performance disks for Coderd, but that provisioners will benefit.	2026-01-06 14:29:41 +04:00
Yevhenii Shcherbina	f792f0b162	docs: introduce landjail for boundary (#21420 )	2026-01-05 18:54:49 -05:00
Asher	4a97df3768	chore: rename flag to disable template insights (#21329 ) Because this affects more than just the template insights page (specifically it also affects the deployment stats endpoint which is shown on bottom bar and Prometheus), the group is being renamed generically to just "stats collection". In the future if we need to affect the other stats we can put those options here. Then, because this change only affects a portion of stats, specifically usage stats like connection and application time, bytes sent, etc, add a new sub-group called "usage stats". Then finally add back the "enable" flag. This also gives us a place to one day place an "anonymize" flag if we need to go that route.	2026-01-05 11:44:06 -09:00
blinkagent[bot]	874f3994b5	docs: update VS Code Web subpath comment to reflect current support (#21375 ) Co-authored-by: blink-so[bot] <211532188+blink-so[bot]@users.noreply.github.com>	2026-01-02 17:16:27 +05:00
Susana Ferreira	b97572285a	feat: add core AI MITM proxy daemon (#21296 ) ## Description Adds the core AI Bridge MITM proxy daemon. This proxy intercepts HTTPS traffic, decrypts it using a configured CA certificate, and forwards requests to AIBridge for processing. ## Changes * Added `aibridgeproxyd` package with the core proxy server implementation * Added configuration options: `CODER_AIBRIDGE_PROXY_ENABLED`, `CODER_AIBRIDGE_PROXY_LISTEN_ADDR`, `CODER_AIBRIDGE_PROXY_CERT_FILE`, `CODER_AIBRIDGE_PROXY_KEY_FILE` * Added tests for server initialization and MITM functionality Closes https://github.com/coder/internal/issues/1180	2025-12-29 15:31:51 +00:00
Danielle Maywood	05529139bc	feat(coderd): support deleting dev containers (#21248 ) Add an endpoint to coderd to support deleting dev containers	2025-12-24 12:34:39 +00:00
Marcin Tojek	0af038bddd	docs: group enumerated values by property in API docs (#21372 ) Fixes #13840	2025-12-22 16:19:25 +01:00
Danielle Maywood	44a46db487	feat(agent): support deleting dev containers (#21247 ) Add logic to the agent, and an endpoint, to allow requesting and then deleting a Dev Container and its related agent.	2025-12-22 11:28:31 +00:00
Rowan Smith	81cbf03a52	chore: fix typo in organization roles create help text (#21352 ) A simple typo fix to the help text stidin > stdin ``` ➜ coder git:(org_role_fix) ✗ coder organizations roles create -h coder v2.29.1+59cdd7e USAGE: coder organizations roles create [flags] <role_name> Create a new organization custom role - Run with an input.json file: $ coder organization -O <organization_name> roles create --stidin < role.json ```	2025-12-22 11:24:00 +11:00
Rowan Smith	0ba3f7e9fd	chore: update organizations.md for Terraform provider support (#21300 )	2025-12-21 06:07:31 +11:00
Bjorn Robertsson	5b3c24c02f	docs: document multiple agents for port-forwarding (#21221 ) Co-authored-by: Atif Ali <atif@coder.com>	2025-12-19 11:45:51 +00:00
Jason Barnett	f9087d6feb	fix: correct Slack webhook example code in documentation (#21295 ) Fixes #21294	2025-12-17 11:27:39 +01:00
blinkagent[bot]	55f4efd011	docs: update Codex CLI compatibility in AI Bridge docs (#21292 ) Co-authored-by: blink-so[bot] <211532188+blink-so[bot]@users.noreply.github.com>	2025-12-16 16:34:48 +00:00
Mathias Fredriksson	dac822b7f4	refactor: remove deprecated AITaskPromptParameterName constant (#21023 ) This removes the deprecated AITaskPromptParameterName constant and all backward compatibility code that was added for v2.28. - Remove AITaskPromptParameterName constant from codersdk/aitasks.go - Remove backward compatibility code in coderd/aitasks.go that populated the "AI Prompt" parameter for templates that defined it - Remove the backward compatibility test (OK AIPromptBackCompat) - Update dbfake to no longer set the AI Prompt parameter - Remove AITaskPromptParameterName from frontend TypeScript types - Remove preset prompt read-only feature from TaskPrompt component - Update docs to reflect that pre-2.28 definition is no longer supported Task prompts are now exclusively stored in the tasks.prompt database column, as introduced in the migration that added the tasks table.	2025-12-16 15:14:59 +00:00
Ethan	42e964ff49	docs: fix typo in MCP documentation (#21287 ) Fix a typo in the MCP documentation where "seems" should be "sees": > These inner loops are not relayed back to the client; all it sees is the result of this loop. Found while reading the docs.	2025-12-16 14:08:21 +05:00
Steven Masley	8fefd91e4a	feat!: support PKCE in the oauth2 client's auth/exchange flow (#21215 ) Breaking Change: Existing oauth apps might now use PKCE. If an unknown IdP type was being used, and it does not support PKCE, it will break. To fix, set the PKCE methods on the external auth to `none` ``` export CODER_EXTERNAL_AUTH_1_PKCE_METHODS=none ```	2025-12-15 17:41:47 +00:00
George K	103967ed02	feat: add sharing info to /workspaces endpoint (#21049 ) closes: https://github.com/coder/internal/issues/858 Similar to https://github.com/coder/coder/pull/19375, this one uses system permissions for fetching actual user and group data. Modifies the `workspaces_expanded` view to fetch the required data; this way it's made available to all code paths that make use of it. Also fixes a bug in a test helper function that can result in `null` being saved to the DB for `user_acl` or `group_acl` and break tests; a defensive check constraint that prevents this is worth a PR, e.g: `ALTER TABLE workspaces ADD CONSTRAINT group_acl_is_object CHECK (jsonb_typeof(group_acl) = 'object');` Also adds missing `OwnerName` in `ConvertWorkspaceRows`.	2025-12-15 08:42:08 -08:00
Asher	27f0413347	feat: add flag to disable template insights (#20940 ) Closes #20399 To summarize the original commit messages: - Do not log stats to the database. - Return errors on the insight endpoints. - Update the frontend to show those errors. - Also fixes an issue with getting the user status count via codersdk, since I added a test to ensure it was not disabled by this flag and it was sending the wrong payload.	2025-12-14 03:00:03 +00:00
Kacper Sawicki	6f86f67754	feat(coderd): add overload protection with rate limiting and concurrency control (#21161 ) ## Summary This adds configurable overload protection to the AI Bridge daemon to prevent the server from being overwhelmed during periods of high load. Partially addresses coder/internal#1153 (rate limits and concurrency control; circuit breakers are deferred to a follow-up). ## New Configuration Options \| Option \| Environment Variable \| Description \| Default \| \|--------\|---------------------\|-------------\|---------\| \| `--aibridge-max-concurrency` \| `CODER_AIBRIDGE_MAX_CONCURRENCY` \| Maximum number of concurrent AI Bridge requests. Set to 0 to disable (unlimited). \| `0` \| \| `--aibridge-rate-limit` \| `CODER_AIBRIDGE_RATE_LIMIT` \| Maximum number of AI Bridge requests per second. Set to 0 to disable rate limiting. \| `0` \| ## Behavior When limits are exceeded: - Concurrency limit: Returns HTTP `503 Service Unavailable` with message "AI Bridge is currently at capacity. Please try again later." - Rate limit: Returns HTTP `429 Too Many Requests` with `Retry-After` header. Both protections are optional and disabled by default (0 values). ## Implementation The overload protection is implemented as reusable middleware in `coderd/httpmw/ratelimit.go`: 1. `RateLimitByAuthToken`: Per-user rate limiting that uses `APITokenFromRequest` to extract the authentication token, with fallback to `X-Api-Key` header for AI provider compatibility (e.g., Anthropic). Falls back to IP-based rate limiting if no token is present. Includes `Retry-After` header for backpressure signaling. 2. `ConcurrencyLimit`: Uses an atomic counter to track in-flight requests and reject when at capacity. The middleware is applied in `enterprise/coderd/aibridge.go` via `r.Group` in the following order: 1. Concurrency check (faster rejection for load shedding) 2. Rate limit check Note: Rate limiting currently applies to all AI Bridge requests, including pass-through requests. Ideally only actual interceptions should count, but this would require changes in the aibridge library. ## Testing Added comprehensive tests for: - Rate limiting by auth token (Bearer token, X-Api-Key, no token fallback to IP) - Different tokens not rate limited against each other - Disabled when limit is zero - Retry-After header is set on 429 responses - Concurrency limiting (allows within limit, rejects over limit, disabled when zero)	2025-12-11 16:38:54 +01:00
Jiachen Jiang	05b02cf887	docs: add deprecation warning to gateway docs and direct to toolbox (#21210 ) See a preview link here: https://coder.com/docs/@gateway-deprecation-docs/user-guides/workspace-access/jetbrains/gateway --------- Co-authored-by: david-fraley <67079030+david-fraley@users.noreply.github.com>	2025-12-10 17:11:25 +00:00
Mathias Fredriksson	8f15caad22	docs: add dev container screenshots (#21191 ) Add screenshots to the dev containers user guide: - Running dev containers with sub-agents (index.md, working-with-dev-containers.md) - Discovered dev containers with Start button (index.md) - Outdated status with rebuild option (working-with-dev-containers.md) - Display apps disabled (customizing-dev-containers.md) Also deletes the outdated devcontainer-agent-ports.png. Refs #21157	2025-12-09 18:08:58 +00:00

1 2 3 4 5 ...

2148 Commits