github docker/docker-agent v1.149.0

6 hours ago

This release adds GitHub repository skill loading, evaluator routing through the models gateway, and several model pricing and capability fixes, along with TUI stability improvements and expanded debug tool hook support.

What's New

  • Adds support for loading skills from public GitHub repositories via https://github.com/owner/repo URLs in skills: entries
  • Runs tool_input_transform and post_tool_use hooks on debug tool calls, completing the full hook pipeline for docker agent debug tool
  • Adds an evaluator judge backend, enabling a dedicated evaluator service as an alternative to the LLM-as-judge path
  • Routes evaluators through the models gateway by default and adds OpenAI Decisions as a native evaluator backend
  • Prices OpenAI fast, priority, and ultrafast service tiers based on the tier actually served in the response rather than the requested tier
  • Adds a Nativ provider guide covering setup, authentication, and API support for the local Apple Silicon inference server

Bug Fixes

  • Fixes ChatGPT models incorrectly treated as text-only by inferring image input capability from the matching OpenAI catalogue entry when no direct ChatGPT entry exists
  • Fixes provider IDs (fireworks-ai, togetherai, moonshotai, opencode) to align with canonical models.dev catalogue IDs, preserving legacy IDs as aliases
  • Fixes custom and gateway provider filters being dropped when resolving model providers
  • Fixes fast/priority and ultrafast service-tier cost multipliers falling back to standard rates when base_url is explicitly set to the official OpenAI endpoint
  • Fixes redundant full re-renders triggered by token usage events when the sidebar is hidden in the TUI
  • Fixes the lean TUI working indicator to remain above the input with a permanent separator, replacing inline spinners with static markers to avoid redraws on every tick

Technical Changes

  • Separates loader defaults from runtime registration so building loader options no longer mutates process-wide runtime state as a side effect
  • Centralizes loaded runtime configuration assembly and narrows the chatserver runtime contract
  • Runs tool_response_transform hooks on debug tool output (previously missing from the debug tool path)
  • Updates documentation for the defuse loop termination argument to reflect the correct invariant
  • Stabilizes the debug-tool test on Windows by removing a flaky native shell dependency from the affected test case

What's Changed

  • chore(deps): bump the actions group across 1 directory with 4 updates by @dependabot[bot] in #4510
  • feat: run tool_input_transform and post_tool_use hooks on debug tool calls by @dgageot in #4518
  • docs: update CHANGELOG.md for v1.148.0 by @docker-read-write[bot] in #4519
  • docs(attachment): state the real termination argument for the defuse loop by @PerryLink in #4372
  • fix(tui): avoid redundant frames for hidden-sidebar usage updates by @dgageot in #4523
  • fix: stabilize debug-tool test against windows shell flake by @dgageot in #4524
  • test(runtime): isolate client fixtures on private HTTP transports by @aheritier in #4526
  • test(eval): avoid writing executable runtimes during parallel tests by @aheritier in #4527
  • fix(chatgpt): infer image input from OpenAI catalogue by @aheritier in #4528
  • feat(eval): add evaluator judge backend by @dgageot in #4531
  • feat(skills): load skills from public GitHub repositories by @dgageot in #4525
  • feat: price OpenAI fast/priority and ultrafast by actual service tier by @dgageot in #4532
  • refactor: separate loader defaults, bootstrap runtime opts, narrow chatserver contract by @dgageot in #4533
  • fix(models): align provider IDs with the catalogue by @aheritier in #4530
  • docs: add Nativ provider guide by @dgageot in #4534
  • feat: route evaluators through the models gateway, add OpenAI Decisions by @dgageot in #4537
  • fix: keep lean TUI working indicator above the input by @rumpl in #4538
  • fix: honor configured OpenAI base_url when pricing fast tiers by @dgageot in #4536

New Contributors

Full Changelog: v1.148.0...v1.149.0

Don't miss a new docker-agent release

NewReleases is sending notifications on new releases.