Commit Graph
16609 Commits
Author SHA1 Message Date
Eugene YurtsevandChester Curme a2ff1bb2ed feat(openai): extract gateway metadata from response headers when available (#39706)
Extracts gateway metadata from response headers when it's included.

---------

Co-authored-by: Chester Curme <chester.curme@gmail.com>
2026-08-17 21:21:17 -04:00
Lingjin 9c21d84bcb fix(core): accept non-dict Mapping values in mustache templates (#39680) 2026-08-17 20:29:33 -04:00
ccurme 4f355f38de docs(core): clarify Runnable pipe coercion [closes #39075] (#39707) 2026-08-18 00:25:17 +00:00
Lingjin 77269bad0b fix(anthropic): exclude sibling directories from grep search scope (#39681) 2026-08-17 20:13:01 -04:00
James Yang 300eb71549 fix(core): finalize chain-group runs on BaseException (#39699) 2026-08-17 19:52:46 -04:00
Eugene Yurtsev 4033a4eb7f chore(core): release 1.5.6 (#39704)
Release 1.5.6
langchain-core==1.5.6
2026-08-17 21:00:09 +00:00
5650448a03 feat(core): incorporate gateway metadata to traces (#39703)
This PR sends gateway metadata information (if present in the client
response) as metadata for the llm invocation.

Requires changes corresponding changes in the ChatModel implementations
(e.g., ChatOpenAI) so gateway metadata is picked up from the gateway
response headers.

---------

Signed-off-by: Eugene Yurtsev <eugene@langchain.dev>
Co-authored-by: Copilot Autofix powered by AI <175728472+Copilot@users.noreply.github.com>
Co-authored-by: copilot-swe-agent[bot] <198982749+Copilot@users.noreply.github.com>
Co-authored-by: eyurtsev <3205522+eyurtsev@users.noreply.github.com>
2026-08-17 20:18:26 +00:00
起岚 3a122c60ce fix: typo (#39684)
This PR fixes the documentation issue reported in #39683: corrects
`capabilites` to `capabilities` in `README.md`.

Fixes #39683

Fixes #

---

Read the full contributing guidelines:
https://docs.langchain.com/oss/python/contributing/overview

> **All contributions must be in English.** See the [language
policy](https://docs.langchain.com/oss/python/contributing/overview#language-policy).

If you paste a large clearly AI generated description here your PR may
be IGNORED or CLOSED!

Thank you for contributing to LangChain! Follow these steps to have your
pull request considered as ready for review.

1. PR title: Should follow the format: TYPE(SCOPE): DESCRIPTION

  - Examples:
    - fix(anthropic): resolve flag parsing error
    - feat(core): add multi-tenant support
    - test(openai): update API usage tests
- Allowed TYPE and SCOPE values:
https://github.com/langchain-ai/langchain/blob/master/.github/workflows/pr_lint.yml#L15-L33

2. PR description:

- Write 1-2 sentences that make the change easy to understand: who
benefits, what problem they had, and how this solves it. Prefer a simple
user story over a long summary.
- The `Fixes #xx` line at the top is **required** for external
contributions — update the issue number and keep the keyword. This links
your PR to the approved issue and auto-closes it on merge.
  - If there are any breaking changes, please clearly describe them.
- If this PR depends on another PR being merged first, please include
"Depends on #PR_NUMBER" in the description.

## Release note

3. Run `make format`, `make lint` and `make test` from the root of the
package(s) you've modified.

  - We will not consider a PR unless these three are passing in CI.

4. How did you verify your code works?

Additional guidelines:

- All external PRs must link to an issue or discussion where a solution
has been approved by a maintainer, and you must be assigned to that
issue. PRs without prior approval will be closed.
- PRs should not touch more than one package unless absolutely
necessary.
- Do not update the `uv.lock` files or add dependencies to
`pyproject.toml` files (even optional ones) unless you have explicit
permission to do so by a maintainer.

## Social handles (optional)

Twitter: @
LinkedIn: https://linkedin.com/in/

Fixes #39683
2026-08-17 15:03:26 -04:00
ccurme 6e2d4f4273 chore(openai): update snapshots (#39657) 2026-08-17 13:23:00 -04:00
langchain-oss-model-profiles[bot]andmdrxy 5327463f5c chore(model-profiles): refresh model profile data (#39692)
Automated refresh of model profile data for all in-monorepo partner
integrations via `langchain-profiles refresh`.

🤖 Generated by the [`refresh_model_profiles`
workflow](https://github.com/langchain-ai/langchain/blob/master/.github/workflows/refresh_model_profiles.yml).

## Summary of changes

**3 added · 0 removed · 9 changed** across 3 provider(s).

<details>
<summary>huggingface</summary>

**➕ 1 added**
- `Qwen/Qwen3.8-2.4T-A95B` — 262,144 ctx, 131,072 out, reasoning, tools

</details>

<details>
<summary>openrouter</summary>

**➕ 1 added**
- `z-ai/glm-5.2:free` — 128,000 ctx, 128,000 out, reasoning

**✏️ 9 changed**
- `deepseek/deepseek-v4-flash-0731`: max input tokens 1,048,576 →
1,310,720
- `deepseek/deepseek-v4-pro`: max output tokens 393,216 → 384,000
- `google/gemma-4-26b-a4b-it`: max output tokens 262,144 → 16,384
- `nvidia/nemotron-3-nano-30b-a3b`: max output tokens 228,000 → 262,144
- `nvidia/nemotron-3.5-lightning`: max output tokens 262,144 → 131,072
- `qwen/qwen3.6-27b`: max output tokens 65,536 → 131,072
- `z-ai/glm-5.2`: max output tokens 128,000 → 262,144
- `~deepseek/deepseek-v4-flash-latest`: max input tokens 1,048,576 →
1,310,720; max output tokens 384,000 → 262,144
- `~moonshotai/kimi-latest`: max output tokens 1,048,576 → 974,842

</details>

<details>
<summary>xai</summary>

**➕ 1 added**
- `grok-imagine-image-2.0` — 8,000 ctx, text+image+pdf in

</details>

Co-authored-by: mdrxy <61371264+mdrxy@users.noreply.github.com>
2026-08-17 09:44:58 -04:00
Anay Garodia 313bc541c0 fix(openai): support o-series models in get_num_tokens_from_messages (#38710) 2026-08-17 09:15:44 -04:00
langchain-oss-model-profiles[bot]andmdrxy 82fd04260c chore(model-profiles): refresh model profile data (#39670)
Automated refresh of model profile data for all in-monorepo partner
integrations via `langchain-profiles refresh`.

🤖 Generated by the [`refresh_model_profiles`
workflow](https://github.com/langchain-ai/langchain/blob/master/.github/workflows/refresh_model_profiles.yml).

## Summary of changes

**3 added · 0 removed · 13 changed** across 2 provider(s).

<details>
<summary>huggingface</summary>

**➕ 1 added**
- `deepseek-ai/DeepSeek-V4-Pro-0813` — 1,000,000 ctx, 384,000 out,
reasoning, tools

</details>

<details>
<summary>openrouter</summary>

**➕ 2 added**
- `dots-studio/dots-3-note-preview:free` — 512,000 ctx, 512,000 out,
text+image in, reasoning, tools
- `qwen/qwen3.8-27b` — 262,144 ctx, 131,072 out, text+image+video in,
reasoning, tools

**✏️ 13 changed**
- `bytedance-seed/seed-2.0-lite`: last updated `2026-03-10` →
`2026-02-14`; display name `Seed-2.0-Lite` → `Seed 2.0 Lite`; release
date `2026-03-10` → `2026-02-14`
- `bytedance-seed/seed-2.0-mini`: last updated `2026-02-26` →
`2026-02-14`; display name `Seed-2.0-Mini` → `Seed 2.0 Mini`; release
date `2026-02-26` → `2026-02-14`
- `deepseek/deepseek-v4-flash`: max output tokens 393,216 → 384,000
- `minimax/minimax-m2-her`: display name `MiniMax M2-her` → `MiniMax-M2
Her`
- `qwen/qwen3-30b-a3b`: max output tokens 16,384 → 8,192
- `qwen/qwen3.5-35b-a3b`: max output tokens 262,144 → 65,536
- `qwen/qwen3.5-397b-a17b`: max output tokens 262,144 → 65,536
- `qwen/qwen3.6-27b`: max output tokens 262,144 → 65,536
- `qwen/qwen3.8-2.4t-a95b`: max input tokens 1,010,000 → 1,048,576
- `z-ai/glm-5`: max output tokens 131,072 → 128,000
- `z-ai/glm-5.1`: max output tokens 131,072 → 128,000
- `z-ai/glm-5.2`: max output tokens 131,072 → 128,000
- `~deepseek/deepseek-v4-flash-latest`: max output tokens 262,144 →
384,000

</details>

Co-authored-by: mdrxy <61371264+mdrxy@users.noreply.github.com>
2026-08-17 00:07:20 -04:00
ccurme 9a58107126 release(openrouter): 0.2.8 (#39658) langchain-openrouter==0.2.8 2026-08-14 15:00:38 -04:00
langchain-oss-model-profiles[bot]andmdrxy c197a7b6f7 chore(model-profiles): refresh model profile data (#39646)
Automated refresh of model profile data for all in-monorepo partner
integrations via `langchain-profiles refresh`.

🤖 Generated by the [`refresh_model_profiles`
workflow](https://github.com/langchain-ai/langchain/blob/master/.github/workflows/refresh_model_profiles.yml).

## Summary of changes

**11 added · 0 removed · 12 changed** across 3 provider(s).

<details>
<summary>fireworks-ai</summary>

**➕ 6 added**
- `accounts/fireworks/models/deepseek-v4-pro-0813` — 1,000,000 ctx,
384,000 out, reasoning, tools
- `accounts/fireworks/models/inkling` — 1,048,576 ctx, 1,048,576 out,
text+image+audio in, reasoning, tools
- `accounts/fireworks/models/muse-glimmer-30b` — 131,072 ctx, 131,072
out, text+image in, reasoning, tools
- `accounts/fireworks/models/nemotron-3-ultra-nvfp4` — 262,144 ctx,
128,000 out, reasoning, tools
- `accounts/fireworks/models/nemotron-lightning-3p5-30b-a3b` — 262,144
ctx, 262,144 out, reasoning, tools
- `accounts/fireworks/models/qwen3p8-max` — 262,144 ctx, 131,072 out,
reasoning, tools

</details>

<details>
<summary>huggingface</summary>

**➕ 4 added**
- `Qwen/Qwen2.5-Coder-32B-Instruct` — 131,072 ctx, 8,192 out, tools
- `Qwen/Qwen3-30B-A3B` — 40,960 ctx, 16,384 out, reasoning, tools
- `deepseek-ai/DeepSeek-V3-0324` — 163,840 ctx, 163,840 out, tools
- `meta-llama/Llama-3.1-8B-Instruct` — 131,072 ctx, 4,096 out, tools

</details>

<details>
<summary>openrouter</summary>

**➕ 1 added**
- `google/gemini-3.7-flash` — 1,048,576 ctx, 65,536 out,
text+image+audio+video+pdf in, reasoning, tools

**✏️ 12 changed**
- `arcee-ai/trinity-large-thinking`: last updated `2026-04-01` →
`2026-05-28`
- `deepseek/deepseek-r1`: max input tokens 163,840 → 64,000
- `deepseek/deepseek-v4-flash-0731`: max output tokens 384,000 → 393,216
- `deepseek/deepseek-v4-pro-0813`: added structured output
- `google/gemini-3.1-flash-lite-image`: max output tokens 65,536 →
66,000
- `liquid/lfm-2.5-2.6b:free`: max output tokens 32,768 → 8,192
- `meta-llama/llama-3.1-8b-instruct`: display name `Llama 3.1 8B
Instruct` → `Llama-3.1-8B-Instruct`
- `nvidia/nemotron-3.5-lightning`: max input tokens 1,048,576 →
1,000,000
- `qwen/qwen3-next-80b-a3b-instruct`: max output tokens 16,384 → 262,144
- `qwen/qwen3-next-80b-a3b-thinking`: max output tokens 262,144 → 32,768
- `qwen/qwen3-vl-30b-a3b-instruct`: max output tokens 16,384 → 32,768
- `qwen/qwen3.8-2.4t-a95b`: max input tokens 1,000,000 → 1,010,000; max
output tokens 52,429 → 262,144

</details>

Co-authored-by: mdrxy <61371264+mdrxy@users.noreply.github.com>
langchain-core==1.5.5
2026-08-14 11:13:52 -07:00
ccurme 555702e1c6 release(core): 1.5.5 (#39655) 2026-08-14 14:13:15 -04:00
ccurme e32fa9a52e chore(infra): add fields for social handles in issue templates (#39654) 2026-08-14 12:16:10 -04:00
ccurme d6cd98a5cc release(openai): 1.5.1 (#39653) langchain-openai==1.5.1 2026-08-14 11:36:30 -04:00
Johannes du Plessis 7b954aa9ef fix(openai): preserve streamed encrypted reasoning (#39635) 2026-08-14 11:33:06 -04:00
ccurme 0082ce0de9 chore(infra): support langsmith gateway in CI (#39651) 2026-08-14 09:28:46 -04:00
pxmps f9ee55d94c fix(core): make abatch_iterate consistent with batch_iterate for None and zero size (#39367) 2026-08-13 17:27:44 -04:00
langchain-oss-model-profiles[bot]andmdrxy 162f9d9582 chore(model-profiles): refresh model profile data (#39625)
Automated refresh of model profile data for all in-monorepo partner
integrations via `langchain-profiles refresh`.

🤖 Generated by the [`refresh_model_profiles`
workflow](https://github.com/langchain-ai/langchain/blob/master/.github/workflows/refresh_model_profiles.yml).

## Summary of changes

**5 added · 1 removed · 3 changed** across 3 provider(s).

<details>
<summary>deepseek</summary>

**✏️ 1 changed**
- `deepseek-v4-pro`: last updated `2026-04-24` → `2026-08-12`; removed
open weights; release date `2026-04-24` → `2026-08-12`

</details>

<details>
<summary>openrouter</summary>

**➕ 4 added**
- `bytedance-seed/seed-2-1-turbo` — 262,144 ctx, 262,144 out,
text+image+video in, reasoning, tools
- `deepseek/deepseek-v4-pro-0813` — 1,048,576 ctx, 384,000 out,
reasoning, tools
- `qwen/qwen3.8-2.4t-a95b` — 1,000,000 ctx, 52,429 out, reasoning, tools
- `x-ai/grok-4.6` — 500,000 ctx, 500,000 out, text+image+pdf in,
reasoning, tools

**➖ 1 removed**
- `inclusionai/ling-3.0-tiny:free`

**✏️ 2 changed**
- `anthracite-org/magnum-v4-72b`: max input tokens 16,384 → 32,768; max
output tokens 2,048 → 4,096
- `nvidia/nemotron-3.5-lightning`: max input tokens 262,144 → 1,048,576;
added tool calling

</details>

<details>
<summary>xai</summary>

**➕ 1 added**
- `grok-4.6` — 500,000 ctx, 500,000 out, text+image+pdf in, reasoning,
tools

</details>

Co-authored-by: mdrxy <61371264+mdrxy@users.noreply.github.com>
2026-08-13 10:02:02 -07:00
ccurme a2f02abedb chore(langchain): update docs on error handling for json schema (#39632) 2026-08-13 15:32:24 +00:00
Hunter Lovell cde529347a fix(core): respect pydantic aliases when validating tool inputs (#39572) 2026-08-13 09:19:50 -04:00
ccurme fb8853d246 release(openai): 1.5.0 (#39629) langchain-openai==1.5.0 2026-08-13 13:00:21 +00:00
ccurmeandMason Daugherty ee88091dc0 feat(openai): support openai 3.0 SDK (#39613)
Co-authored-by: Mason Daugherty <github@mdrxy.com>
2026-08-13 12:44:32 +00:00
Mason Daugherty c5b2f95b5e release(anthropic): 1.5.6 (#39622)
Bumps `langchain-anthropic` from 1.5.5 to 1.5.6, a patch release
containing two fixes since the last release:

- `fix(anthropic):` normalize `tool_search_tool_result` blocks (#39621)
- `fix(anthropic):` correct model profile data for Fable 5, Sonnet 5,
Opus 4.1 (#39604)
langchain-anthropic==1.5.6
2026-08-12 18:17:55 -07:00
Hunter Lovell 3c6c3ba7bf fix(anthropic): normalize tool_search_tool_result blocks (#39621)
fixes #37584

* `tool_search_tool_result` blocks were being passed through unchanged,
which when streamed included a disallowed `index` field (this was
causing 400's with `ProviderToolSearchMiddleware`)
* this adds an additional case in the content block normalization step
to make sure that doesn't get tracked

ref:
https://platform.claude.com/docs/en/api/messages/create#tool_search_tool_result_block_param
2026-08-12 18:09:24 -07:00
Nishitha Mandopen-swe[bot] a648460cfd fix(core): issues in merging chunks (#39535)
Closes #38064, #35259, #38850

`merge_dicts`, `merge_lists`, and `AddableDict` all guessed at merge
semantics for streaming chunks in ways that silently corrupted data
instead of failing loudly:

- `merge_dicts` fell into the `int` branch for differing `bool` values
(since `bool` subclasses `int`) and summed them, turning `True + False`
into `1`. It now raises `TypeError`, consistent with other unmergeable
types.
- `merge_lists` used `"index" in e_left` on untyped list elements; when
an element was a plain `str` containing the literal substring `"index"`,
it then subscripted the string and raised an unrelated `TypeError`. It
now checks `isinstance(e_left, dict)` first.
- `AddableDict.__add__`/`__radd__` silently discarded the left-hand
value on any `TypeError` from `chunk[key] + other[key]`. They now
re-raise a `TypeError` naming the key and both types.

### Release note

`merge_dicts` now raises `TypeError` for differing boolean values at the
same key instead of silently summing them to an `int`; `AddableDict`
addition now raises `TypeError` on type-incompatible keys instead of
silently dropping data; `merge_lists` no longer misidentifies non-dict
elements as index-keyed.

Co-authored-by: open-swe[bot] <open-swe@users.noreply.github.com>
2026-08-12 17:40:35 -04:00
Hunter Lovell 13b2f8d727 fix(core): handle v1 base model validation in async path (#39576) 2026-08-12 15:49:24 -04:00
ccurme 7fb045b744 fix(fireworks): avoid repeated service_tier [closes #39619] (#39620) 2026-08-12 19:41:51 +00:00
Hunter Lovelland王俊锋 624fd031c2 fix(core): handle tool descriptions for infer_schema=False (#39573)
Co-authored-by: 王俊锋 <37211900+kinch-tech@users.noreply.github.com>
2026-08-12 15:22:04 -04:00
Hunter Lovell dff3b73287 fix(langchain): preserve final repeated schema ordering (#39284) 2026-08-12 15:13:20 -04:00
iroiro147andChester Curme 1a2585ee5f fix(openrouter): preserve cost metadata in usage chunks (#39338)
Co-authored-by: Chester Curme <chester.curme@gmail.com>
2026-08-12 13:39:37 -04:00
ccurmeandSolaris-star 0a86934ab5 fix(core): clear usage metadata callback on exceptions in context manager (#39616)
Co-authored-by: Solaris-star <820622658@qq.com>
2026-08-12 13:32:51 -04:00
Mason Daugherty 2d47a5f398 chore(partners): bump langgraph floor in openai and huggingface lockfiles (#39617) 2026-08-12 09:58:11 -07:00
langchain-oss-model-profiles[bot]andmdrxy 4545c74216 chore(model-profiles): refresh model profile data (#39609)
Automated refresh of model profile data for all in-monorepo partner
integrations via `langchain-profiles refresh`.

🤖 Generated by the [`refresh_model_profiles`
workflow](https://github.com/langchain-ai/langchain/blob/master/.github/workflows/refresh_model_profiles.yml).

## Summary of changes

**4 added · 0 removed · 2 changed** across 1 provider(s).

### openrouter

**➕ 4 added**
- `bytedance-seed/seed-2.0-code` — 262,144 ctx, 131,072 out,
text+image+video in, reasoning, tools
- `liquid/lfm-2.5-2.6b:free` — 128,000 ctx, 32,768 out, reasoning, tools
- `nvidia/nemotron-3.5-lightning` — 262,144 ctx, 262,144 out, reasoning
- `nvidia/nemotron-3.5-lightning:free` — 1,000,000 ctx, 65,536 out,
reasoning, tools

**✏️ 2 changed**
- `z-ai/glm-5.2`: max output tokens 262,144 → 131,072
- `~deepseek/deepseek-v4-flash-latest`: max output tokens 131,072 →
262,144

Co-authored-by: mdrxy <61371264+mdrxy@users.noreply.github.com>
2026-08-12 09:56:35 -07:00
Hunter Lovell b6eccc5f97 fix(core): handle falsy LLM and chat model caches (#39283) 2026-08-12 12:39:16 -04:00
ccurme 5a28e17fbb chore(core): add httpx as an explicit dep (#39612) 2026-08-12 10:11:28 -04:00
Mason Daugherty 0727992d4a fix(anthropic): correct model profile data for Fable 5, Sonnet 5, Opus 4.1 (#39604)
Fixed `ChatAnthropic` model profiles: `claude-fable-5` and
`claude-sonnet-5` now correctly report structured output support,
`claude-fable-5` reports its reasoning effort levels and default, and
the retired `claude-opus-4-1` entry was removed.

---

Users calling `model.profile` on `ChatAnthropic` get wrong capability
data for three models:

- `claude-fable-5` reported `structured_output: false` and no
reasoning-effort metadata, even though Fable 5 supports structured
outputs and all five effort levels (`low` through `max`, defaulting to
`high`).
- `claude-sonnet-5` reported `structured_output: false` despite
supporting structured outputs.
- `claude-opus-4-1` lingered as a sparse augmentation-only entry (six
fields, no token limits or modalities) after models.dev dropped it —
Anthropic retired the model on 2026-08-05, so the entry is removed
rather than backfilled.

Corrections verified against Anthropic's [structured outputs
compatibility
list](https://platform.claude.com/docs/en/build-with-claude/structured-outputs),
[effort
docs](https://platform.claude.com/docs/en/build-with-claude/effort), and
[deprecation
schedule](https://platform.claude.com/docs/en/about-claude/model-deprecations).
`_profiles.py` regenerated with `langchain-profiles refresh --provider
anthropic`.
2026-08-11 21:45:00 -07:00
zerafachris 7cdf9658f0 fix(core): preserve non-str/non-dict items in DictPromptTemplate list values (#39588) 2026-08-11 21:32:33 -04:00
John Kennedyandlangsmith-fleet[bot] 2c3b11c6c4 chore: bump setuptools in Hugging Face lockfile (#39603)
## Summary

- bumps `setuptools` from 81.0.0 to 84.0.0 in
`libs/partners/huggingface/uv.lock`
- resolves Dependabot alert #3955 / GHSA-h35f-9h28-mq5c / CVE-2026-59890
(first patched version: 83.0.0)
- keeps the change scoped to the affected lockfile entry

## Validation

- [x] `uv sync --project libs/partners/huggingface --frozen
--no-install-project --no-dev`
- [x] verified resolved `setuptools==84.0.0` is outside the vulnerable
range
- [x] `git diff --check`

## Notes

`uv lock --locked` reports that the lockfile needs unrelated workspace
metadata updates (`langchain`, `langchain-core`, and `langgraph`). Those
unrelated refreshes were intentionally excluded to keep this security
patch narrow.

Co-authored-by: langsmith-fleet[bot] <langsmith-fleet[bot]@users.noreply.github.com>
2026-08-11 16:49:16 -07:00
Rin 61c5678835 fix(core): raise ValueError when explicit tool_outputs length mismatches tool_calls in tool_example_to_messages (#39142) 2026-08-11 18:54:59 -04:00
Willow Lopez e34cd76346 fix(core): guard malformed Anthropic content blocks (#38670) 2026-08-11 18:50:39 -04:00
ccurme 119bf69a1e release(anthropic): 1.5.5 (#39597) langchain-anthropic==1.5.5 2026-08-11 15:12:52 -04:00
ccurme f4bc5031db release(langchain): 1.3.15 (#39595) langchain==1.3.15 2026-08-11 14:39:46 -04:00
ccurme 5ff19c613a release(core): 1.5.4 (#39592) langchain-core==1.5.4 2026-08-11 13:43:45 -04:00
ccurme 1925966dc6 feat(langchain): expose trace_policy on AgentMiddleware (#38910) 2026-08-11 13:33:54 -04:00
langchain-oss-model-profiles[bot]andmdrxy ce8e8bd8b1 chore(model-profiles): refresh model profile data (#39579)
Automated refresh of model profile data for all in-monorepo partner
integrations via `langchain-profiles refresh`.

🤖 Generated by the [`refresh_model_profiles`
workflow](https://github.com/langchain-ai/langchain/blob/master/.github/workflows/refresh_model_profiles.yml).

## Summary of changes

**3 added · 1 removed · 8 changed** across 1 provider(s).

### openrouter

**➕ 3 added**
- `meta/muse-glimmer-30b` — 131,072 ctx, 131,072 out, text+image in,
reasoning, tools
- `sakana/sakana-namazu` — 262,144 ctx, 65,536 out, text+image+pdf in,
reasoning, tools
- `upstage/solar-pro4` — 524,288 ctx, 131,072 out, reasoning, tools

**➖ 1 removed**
- `openai/gpt-5.3-chat`

**✏️ 8 changed**
- `deepseek/deepseek-v4-pro`: max output tokens 384,000 → 393,216
- `inclusionai/ling-3.0-tiny:free`: added open weights
- `moonshotai/kimi-k3`: added video input
- `openai/gpt-5.2-chat`: max output tokens 16,384 → 32,000
- `qwen/qwen3-coder-30b-a3b-instruct`: max output tokens 32,768 →
262,144
- `qwen/qwen3-next-80b-a3b-thinking`: max output tokens 32,768 → 262,144
- `qwen/qwen3.5-397b-a17b`: max output tokens 65,536 → 262,144
- `~moonshotai/kimi-latest`: added video input

Co-authored-by: mdrxy <61371264+mdrxy@users.noreply.github.com>
2026-08-11 08:49:43 -07:00
ccurmeandPhilemon Schöpf a2a9b1bde4 fix(anthropic): report reasoning tokens in usage metadata (#39590)
Co-authored-by: Philemon Schöpf <philemon.schoepf@otera.ai>
2026-08-11 11:42:29 -04:00
ccurme 39b4e0f9a8 chore(langchain): fix type errors in tests (#39589) 2026-08-11 11:25:58 -04:00