[Misc]feat: adapt to vLLM main (c7aa186d...2a16ece2) - #240
Draft
Meihan-chen wants to merge 41 commits into
Draft
Conversation
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <jcccx.cmh@gmail.com>
Signed-off-by: Meihan-chen <239345480+Meihan-chen@users.noreply.github.com>
Signed-off-by: Meihan-chen <239345480+Meihan-chen@users.noreply.github.com>
Signed-off-by: Meihan-chen <239345480+Meihan-chen@users.noreply.github.com>
Signed-off-by: Meihan-chen <239345480+Meihan-chen@users.noreply.github.com>
Signed-off-by: Meihan-chen <239345480+Meihan-chen@users.noreply.github.com>
Signed-off-by: Meihan-chen <239345480+Meihan-chen@users.noreply.github.com>
Upstream range: c7aa186d67b6f051680831418e957c67f34ba7a2..16e336491e96041946e2cd379a80034fcac439bd Changes absorbed: - mistral_common version bump to 1.11.2 - Mistral tokenizer refactoring (_prepare_apply_chat_template_tools_and_messages -> _validate_apply_chat_template_args) No vllm-ascend code adaptation required - vllm-ascend does not use the changed Mistral tokenizer functions directly. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com> Signed-off-by: main2main-bot <main2main-bot@users.noreply.github.com>
Upstream range: 16e336491e96041946e2cd379a80034fcac439bd..5d0fd87038b123a11dd9c85a05ce2d258e27ce7b Changes absorbed: - Bugfix: PP sampled-token receive condition changed from world_size > 1 to not is_last_rank - Spec decode: _raise_if_multimodal renamed to _warn_if_multimodal for multimodal support with warning - CPU/RISC-V OMP multiprocessing changes (no Ascend impact) - Various CPU/ROCm kernel updates (no Ascend impact) - Compilation codegen interface changes (no Ascend impact - xlite doesn't use generate_execution_code) - Config rope_type backwards compatibility patch (addition only) vllm-ascend adaptations: - model_runner_v1.py: Updated PP sampled-token receive condition - dflash_proposer.py: Renamed _raise_if_multimodal to _warn_if_multimodal Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com> Signed-off-by: main2main-bot <main2main-bot@users.noreply.github.com>
Upstream range: 5d0fd87038b123a11dd9c85a05ce2d258e27ce7b..242afc6bf40d6d088b7b97eda1dd7f00ddcdfa21 Changes absorbed: - [MM][Gemma4] Respect max_soft_tokens in encoder budget (#41799) No vllm-ascend code adaptation required - vllm-ascend does not have Gemma4 model implementation. Commit reference only. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com> Signed-off-by: main2main-bot <main2main-bot@users.noreply.github.com>
Meihan-chen
force-pushed
the
main
branch
2 times, most recently
from
May 20, 2026 03:18
41aae44 to
9af7666
Compare
zhangxinyuehfad
force-pushed
the
main
branch
3 times, most recently
from
June 9, 2026 06:11
e4d6423 to
9f88e4b
Compare
MrZ20
force-pushed
the
main
branch
2 times, most recently
from
June 10, 2026 03:25
677396a to
9554149
Compare
zhangxinyuehfad
force-pushed
the
main
branch
5 times, most recently
from
June 16, 2026 04:48
c14c88e to
9768184
Compare
zhangxinyuehfad
force-pushed
the
main
branch
9 times, most recently
from
June 18, 2026 04:04
de24f70 to
60cd7c9
Compare
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Automated adaptation to upstream vLLM main branch changes.
Commit range: c7aa186d67b6f051680831418e957c67f34ba7a2...2a16ece2d342c0c154a4949ad317b521f8c04ec4
Created Commits
59076f5d1ca9b73af617ab34d2e458080268f529main2main: update vLLM commit reference to 16e33649917370e89373ea551fcebadd3d8d1359c02a6e8emain2main: update vLLM commit reference to 5d0fd87cebe1f2e9855c2b2300a3626f765c37fe3733672main2main: update vLLM commit reference to 242afc6580e7b0bf4ef575938358bcb30fd6b10ea6f466cmain2main: update vLLM commit reference to df8e63fafca80f716b4a63278bc89c440e22383ed1cbc7cmain2main: update vLLM commit reference to f39bcf1fd4140f85413df0f0626ceaf583476910bc54477main2main: update vLLM commit reference to 27e00574a7858b883429f9197a0b6678aec790ccc2bd79emain2main: update vLLM commit reference to 38e1667d0a03e142cb62a5926b534979aaa88e8f7b73424main2main: update vLLM commit reference to ca3e62d09e37f9d1a74895b0612d0be8994889a86e929e4main2main: update vLLM commit reference to 20cac2664317d20e0ec625cf7418deee3ba8ee6385a09c4main2main: update vLLM commit reference to 51f22dc19d4702441c486368529a16e65e3f5b4e2b57248main2main: update vLLM commit reference to 9c0812fbd4b616fcade0f876626c177b3eef4f235ef22d3main2main: update vLLM commit reference to ffee741d775b2abf216bfeee9afebfc2bcd54a1baefb4d4main2main: update vLLM commit reference to b3945cc341a48e46bb18a5b9c4c0f6b9a15c0acad71810emain2main: update vLLM commit reference to 2a16ece (target)main2main Summary
Main2Main Summary
Status: completed
Upstream range: c7aa186d67b6f051680831418e957c67f34ba7a2..2a16ece2d342c0c154a4949ad317b521f8c04ec4
Reached upstream commit: 2a16ece2d342c0c154a4949ad317b521f8c04ec4
Steps: 14/14
CI suite: e2e-main2main (not run - Ascend NPU hardware unavailable)
Result
Successfully absorbed 50 upstream commits across 14 steps. The target commit was fully reached with minimal code adaptations required. Most upstream changes were platform-specific (CPU, ROCm, XPU) or feature additions that did not affect vllm-ascend's Ascend-specific implementation.
Completed Steps
Changes Made
CI Verification
Adaptation Summary
Only 2 of 14 steps required code adaptation:
Step-2: Bugfix for PP sampled-token receive condition in model_runner_v1.py. The upstream changed
get_pp_group().world_size > 1tonot get_pp_group().is_last_rankto correctly receive sampled tokens only on non-last ranks.Step-2: Speculative decoding method rename in dflash_proposer.py. The upstream changed
_raise_if_multimodalto_warn_if_multimodalto allow multimodal models with a warning instead of raising an error.All other steps were commit reference updates only because:
Pre-Completion Checklist
Notes
This run was executed in an environment without Ascend NPU hardware, so full CI verification could not be performed.
The pipeline should be verified by running e2e-main2main suite on Ascend hardware before merging to main.
No pushes or PRs were created as per the task requirements.
vLLM version: v0.20.1
vLLM main: vllm-project/vllm@c7aa186