Commit Graph

4547 Commits

Author SHA1 Message Date
Saryev Rustam 7b52cef2e6 feat(ai): add OpenRouter OAuth support (#6927)
Closes #6814
2026-07-22 15:48:39 +02:00
David Brailovsky fe42ba5b38 write tui debug/crash logs into the configured pi agent dir (#6958)
fixes #6652
2026-07-22 15:48:25 +02:00
Mario Zechner 346c85473d Merge branch 'agent-harness-tools' 2026-07-22 15:36:28 +02:00
Mario Zechner e32c1491b5 feat(agent): align harness execution tools 2026-07-22 15:36:06 +02:00
Armin Ronacher 5dc40fee33 feat(ai): derive generated model types from JSON 2026-07-22 13:13:36 +02:00
David Brailovsky 906b40a753 Merge pull request #6955 from earendil-works/fix/handle-openai-previous-response-not-found-error
handle openai websocket previous_response_not_found error
2026-07-22 12:41:58 +02:00
David Brailovsky c5dcb26000 handle openai websocket previous_response_not_found error
fixes #6931
2026-07-22 10:26:24 +00:00
Mario Zechner 2d0b2294cd feat(agent): merge main into agent-harness-tools 2026-07-22 11:56:35 +02:00
Matthew bc41f612da fix(ai): cache OpenRouter tool results
closes #6940
2026-07-22 11:48:33 +02:00
Mario Zechner 6f95a33888 fix(ai): enable cache control for OpenRouter aliases
closes #6940
2026-07-22 11:46:21 +02:00
Christian Klotz 33e40c3e15 fix(ai): retry DNS transport failures (#6946)
Fixes #6904
2026-07-22 11:26:19 +02:00
zay a5afc3f171 feat(ai): add Kimi Code subscription OAuth login (#6935)
Add a device authorization grant flow (RFC 8628) for the kimi-coding
provider, mirroring the official Kimi Code CLI: device authorization and
token polling against https://auth.kimi.com, refresh-token rotation with
retry/backoff, and Bearer auth via toAuth headers. The provider now
offers 'Sign in with Kimi Code' alongside the existing KIMI_API_KEY
auth. Host is overridable via KIMI_CODE_OAUTH_HOST / KIMI_OAUTH_HOST.

Co-authored-by: Mario Zechner <badlogicgames@gmail.com>
2026-07-22 08:35:03 +02:00
David Brailovsky 1ae064099c generate-models: use reasoning options from models.dev (#6928)
* generate-models: use reasoning options from models.dev

* fix tests
2026-07-22 08:33:34 +02:00
Mario Zechner dd6bea41ef Add [Unreleased] section for next cycle 2026-07-21 18:40:11 +02:00
Mario Zechner 20be4b18d4 Release v0.81.1 2026-07-21 18:40:08 +02:00
Mario Zechner b9e5c5d941 fix(agent): restore streamFn extension compatibility
Keep streamFn required for typed callers while preserving the legacy runtime fallback for extensions that omit it.\n\nfixes #6915
2026-07-21 18:37:07 +02:00
Armin Ronacher b142504125 fix(coding-agent): defer catalog refresh until after TUI startup 2026-07-21 18:25:38 +02:00
Armin Ronacher a82289637b feat(ai): add new Gemini generated models 2026-07-21 18:25:27 +02:00
David Brailovsky 7540da4016 add docs for new event types 2026-07-21 15:40:11 +00:00
David Brailovsky 243f64be59 report aborted retry attempts as unsuccessful 2026-07-21 15:34:31 +00:00
David Brailovsky ceac988635 Merge remote-tracking branch 'origin/main' into fix/issue-6647-retry-summary-requests-2 2026-07-21 17:15:54 +02:00
Mario Zechner 37eb243d26 feat(agent): add AgentHarness execution tools 2026-07-21 17:03:28 +02:00
David Brailovsky 959cc1897e fix moonshot/kimi3 compat properties 2026-07-21 16:49:36 +02:00
Cristina Poncela Cubeiro 899fe26911 fix: flaky test (#6905) 2026-07-21 15:53:51 +02:00
Mario Zechner 3dcb411e89 Add [Unreleased] section for next cycle 2026-07-21 15:28:05 +02:00
Mario Zechner 9c480b6ad2 Release v0.81.0 2026-07-21 15:28:05 +02:00
Mario Zechner d35237cfda fix: build packages before release tests 2026-07-21 14:51:22 +02:00
Mario Zechner c23b58a6d8 feat(ai): update generated image models 2026-07-21 14:51:14 +02:00
Mario Zechner c901d9a877 docs: audit unreleased changelogs 2026-07-21 14:47:00 +02:00
Mario Zechner 864b35c462 fix(coding-agent): show llama download progress 2026-07-21 14:25:10 +02:00
Mario Zechner 109c4125d2 feat(agent): prepare SQLite storage package for publishing 2026-07-21 14:13:14 +02:00
David Brailovsky 162179af5a add branch/compact retries to agent-harness 2026-07-21 12:11:39 +00:00
Cristina Poncela Cubeiro 8495f9d0d6 chore: rename orchestrator to server (#6898) 2026-07-21 13:32:42 +02:00
David Brailovsky 8e53e0e49c compaction & branch summarization follow retry policy
fixes #6647

compaction (auto & manual) and branch summarization retry on transient failures.
use the same retry policy from settings.
emit events for the tui to show indication of retries
2026-07-21 10:07:20 +00:00
Cristina Poncela Cubeiro 9e7582aa03 feat: sqlite session storage (#6594)
This PR:

- Adds retainedTail to compaction entries in the new agent harness so we don't have to walk up the tree for the 2000 tokens before compaction,
- Changes getPathToRoot to getPathToRootOrCompaction to only load until last compaction, as unnecessary to access all nodes where it is called,
- Adds a SQLite storage backend, in a separate packages/session-backend-sqlite, with a migration system and schemas as per on-site discussions: sessions to match session header messages (except for metadata, which I couldn't understand what it's used for or where it gets written, so I omitted it), session_entries for shared entry types as columns plus payload as a json for what remains, session_sequences to represent the append-only, serialized nature of the jsonl files, branch_entries to attribute nodes to branches (relationship one-to-many), and session_materialized with the session info (see /session in TUI) to act as a "cache" or quick-access for costs, message count, token info, labels, session name, and model-thinking-level config (e.g. for fast resume).
- This is compatible with the new agent harness Session abstraction.
2026-07-21 11:36:31 +02:00
Armin Ronacher 54fad505b9 fix(coding-agent): prefer newer generated model catalogs 2026-07-21 11:11:08 +02:00
David Brailovsky 890b3547af run generate-models
fixes #6891
2026-07-21 07:48:27 +00:00
David Brailovsky 31dc078bf8 update brace-expansion version 2026-07-21 06:07:44 +00:00
Armin Ronacher c8c3cd499f fix(ai): validate generated model data before builds 2026-07-20 22:29:06 +02:00
Mario Zechner 1235c0ec64 fix(agent): decouple agent streams from compat
closes #6851
2026-07-20 17:54:36 +02:00
Armin Ronacher 3a40794ea1 feat(ai): regenerate models 2026-07-20 17:19:40 +02:00
Mario Zechner c889eb8809 fix(coding-agent): defer startup model catalog refresh 2026-07-20 17:13:06 +02:00
Mario Zechner f8b74a4507 fix: complete extension usage accounting
closes #6509
2026-07-20 17:06:22 +02:00
David Brailovsky 2fd3868401 add usage info to branch summary, compaction and tool result entries (#6671)
* add usage info to branch summary entries

* add usage to compaction entries

* allow custom tools to report llm usage in tool results

* allow observing and patching usage in tool_result hooks

* agent-harness: save usage in entries

for compaction, branch summaries and tool results
2026-07-20 16:41:43 +02:00
Cristina Poncela Cubeiro c179395218 feat: get_available_thinking_levels rpc (#6865) 2026-07-20 16:29:10 +02:00
Cristina Poncela Cubeiro 1942b2600f fix: env section ignored (#6864) 2026-07-20 16:20:49 +02:00
Mario Zechner 13437ca828 fix(ai): normalize Kimi K2.7 to the canonical coding model
Treat the models.dev k2p7 entry as an alias for kimi-for-coding and regenerate provider catalogs. This restores consistency between generated model types and values and unblocks repository type checks.
2026-07-20 14:04:21 +02:00
Mario Zechner 3595e080cb fix(tui): keep paste registry in sync when deleting paste markers
Undo snapshots now restore paste content and counters alongside editor text. Paste marker renumbering shifts registry entries in ascending ID order before rewriting markers, preventing literal or incorrect paste content on submit.\n\nCloses #6844
2026-07-20 14:04:21 +02:00
QuintinShaw bbb91fa8ae feat(ai): add Qwen Token Plan as built-in provider (#6858)
Add Alibaba Cloud Model Studio Token Plan subscription service as two
built-in API-key providers: qwen-token-plan (international, Singapore)
and qwen-token-plan-cn (China, Beijing).

Each provider exposes 15 text-generation models (Qwen, DeepSeek, GLM,
Kimi, MiniMax) via the OpenAI-compatible endpoint with DashScope
enable_thinking support. Model metadata is sourced from models.dev;
qwen3.8-max-preview is hardcoded until models.dev includes it.

Also fixes kimi-coding test references (k2p7 -> kimi-for-coding) after
models.dev catalog update picked up by generate-models.

Closes #6850
2026-07-20 13:53:30 +02:00
Cristina Poncela Cubeiro d9f7f81473 fix: tool_call_id error when switching (#6854) 2026-07-20 13:51:16 +02:00