zhongdaor-nv
|
49eca14b69
|
fix: optimize uuid calculation (#6596)
Signed-off-by: zhongdaor <zhongdaor@nvidia.com>
|
2026-02-25 15:25:48 -08:00 |
Yuewei Na
|
8483e4a07f
|
chore: Upgrade to tensorrt-llm==1.3.0rc5 (#6579)
Signed-off-by: Yuewei Na <nv-yna@users.noreply.github.com>
Co-authored-by: Yuewei Na <nv-yna@users.noreply.github.com>
|
2026-02-25 12:22:58 -08:00 |
ishandhanani
|
6642e23e0f
|
feat: sglang to 0.5.9 + updated docs (#6518)
Co-authored-by: baihuitian <baihuitian.bht@gmail.com>
Co-authored-by: Cursor <cursoragent@cursor.com>
|
2026-02-24 21:48:25 +00:00 |
Yuewei Na
|
9a15730a7f
|
chore: Upgrade to tensorrt-llm==1.3.0rc3 (#6402)
Signed-off-by: Yuewei Na <nv-yna@users.noreply.github.com>
Co-authored-by: Yuewei Na <nv-yna@users.noreply.github.com>
Co-authored-by: dagil-nvidia <dagil@nvidia.com>
|
2026-02-19 14:12:32 -05:00 |
Qi Wang
|
0abebe3884
|
ci: add known_third_party to avoid pre-commit isort failure (#6292)
|
2026-02-13 19:19:47 -08:00 |
Alec
|
2c6e6d2204
|
chore(deps): bump vLLM to 0.15.1 (#6102)
Signed-off-by: alec-flowers <aflowers@nvidia.com>
Signed-off-by: Alec <35311602+alec-flowers@users.noreply.github.com>
Co-authored-by: Claude Sonnet 4.5 <noreply@anthropic.com>
|
2026-02-11 21:19:31 +00:00 |
Dmitry Tokarev
|
4220771fba
|
fix: Cleanup pytest markers, enable gpu_0 tests on trtllm arm, reduce log noise (#6124)
Signed-off-by: Dmitry Tokarev <dtokarev@nvidia.com>
|
2026-02-11 16:00:58 -05:00 |
Qi Wang
|
ae09b92916
|
ci: apply static type check to vllm multimodal handlers (#6027)
|
2026-02-10 10:48:30 -08:00 |
Tushar Sharma
|
6401e34daf
|
ci: Transition deploy tests to pytest framework (#5874)
Signed-off-by: Tushar Sharma <tusharma@nvidia.com>
|
2026-02-06 21:09:34 +00:00 |
Qi Wang
|
f6dd90474b
|
feat: enable go-to-definition for dynamo.runtime, dynamo.nixl and external dependencies (#6026)
|
2026-02-06 12:01:29 -08:00 |
Ayush Agarwal
|
76e0e2076b
|
feat: basic vllm omni pipeline support (#5608)
Signed-off-by: ayushag <ayushag@nvidia.com>
|
2026-02-04 18:35:41 +00:00 |
Konrad Nowicki
|
7e970d44c9
|
feat: image diffusion with SGLang diffusion (#5609)
Signed-off-by: Konrad Nowicki <knowicki@nvidia.com>
Co-authored-by: dagil-nvidia <dagil@nvidia.com>
|
2026-02-03 21:19:21 +00:00 |
Keiven C
|
44986bf53a
|
test: remove hardcoded ports and add timeouts to KVBM tests (#5855)
Signed-off-by: Keiven Chang <keivenc@nvidia.com>
Co-authored-by: Keiven Chang <keivenchang@users.noreply.github.com>
|
2026-02-03 12:21:19 -08:00 |
Anant Sharma
|
73597110ea
|
fix: add msgpack dependency to trtllm section (#5799)
Signed-off-by: Anant Sharma <anants@nvidia.com>
|
2026-01-30 17:29:56 +00:00 |
Tanmay Verma
|
ba711cc1ac
|
chore: Upgrade to Tensorrt-LLM 1.3.0rc1 (#5700)
Co-authored-by: Pavithra Vijayakrishnan <160681768+pvijayakrish@users.noreply.github.com>
|
2026-01-29 15:20:58 -08:00 |
Alec
|
63bee8b303
|
chore: bump vLLM to 0.14.1 (#5691)
Signed-off-by: alec-flowers <aflowers@nvidia.com>
|
2026-01-27 19:51:29 -08:00 |
Pavithra Vijayakrishnan
|
66c369964a
|
chore: version update for 0.9.0 (#5661)
Signed-off-by: pvijayakrish <pvijayakrish@nvidia.com>
|
2026-01-26 17:18:05 -08:00 |
ishandhanani
|
4f981e282a
|
feat: sglang update to 0.5.8 (#5655)
|
2026-01-26 23:46:05 +00:00 |
Aditya Shrish Puranik
|
76fbbc8fab
|
fix: Add cupy-cuda12x to sglang extras (#5627)
Signed-off-by: Aditya Puranik <adityapuranik99@gmail.com>
Co-authored-by: ishandhanani <82981111+ishandhanani@users.noreply.github.com>
|
2026-01-26 13:33:14 -08:00 |
Alec
|
50f1e0e17a
|
chore(deps): bump vLLM to 0.14.0 (#5593)
Signed-off-by: alec-flowers <aflowers@nvidia.com>
Signed-off-by: Alec Flowers <aflowers@nvidia.com>
Signed-off-by: Alec <35311602+alec-flowers@users.noreply.github.com>
Co-authored-by: Claude Opus 4.5 <noreply@anthropic.com>
|
2026-01-24 01:19:06 +00:00 |
jthomson04
|
64ba7dd068
|
chore: Bump TRTLLM to 1.2.0rc6.post2 (#5580)
Signed-off-by: jthomson04 <jwillthomson19@gmail.com>
|
2026-01-22 20:23:40 -08:00 |
Nate Mailhot
|
6dba119d7a
|
chore: update nixl to 0.9.0 (#5528)
|
2026-01-22 15:25:03 -08:00 |
Olga Andreeva
|
224f63f56a
|
test: Fixing kvbm tests/ adding concurrency test (#5474)
|
2026-01-21 14:29:09 -07:00 |
dagil-nvidia
|
6fe88553c5
|
fix: ignore torchao SyntaxWarning in pytest filterwarnings (#5429)
Signed-off-by: Dan Gil <dagil@nvidia.com>
|
2026-01-14 23:32:52 +08:00 |
jthomson04
|
9ca2923de7
|
chore: Revert KVBM v2 transfer integration (#5406)
|
2026-01-13 14:18:07 -08:00 |
Qi Wang
|
4be72b6f60
|
fix: test_consolidator_router_e2e trtllm import failure (#5404)
|
2026-01-13 21:22:07 +00:00 |
Elias Bermudez
|
c3dc3de48d
|
chore: Update aiperf pinned version (#5331)
Signed-off-by: Anant Sharma <anants@nvidia.com>
Co-authored-by: Anant Sharma <anants@nvidia.com>
|
2026-01-13 12:34:21 -05:00 |
Tanmay Verma
|
abd4b5d9a0
|
chore: Upgrade to tensorrt_llm==1.2.0rc6.post1 (#5356)
|
2026-01-12 18:58:00 +00:00 |
ishandhanani
|
92748c937f
|
feat: sglang update to 0.5.7 (#5148)
|
2026-01-08 17:00:03 +08:00 |
Alec
|
a1333a8ddb
|
fix: update vLLM to 0.13.0 with API compatibility fixes (#5222)
Signed-off-by: Vasilis Vagias <vvagias@nvidia.com>
Signed-off-by: alec-flowers <aflowers@nvidia.com>
Co-authored-by: Vasilis Vagias <vvagias@nvidia.com>
|
2026-01-06 15:08:08 -08:00 |
Tushar Sharma
|
cf433e6825
|
chore: update all copyright headers in repo to 2026 (#5130)
Signed-off-by: Tushar Sharma <tusharma@nvidia.com>
|
2026-01-02 22:08:23 +00:00 |
Nate Mailhot
|
3c7ed61d05
|
chore: TRTLLM 1.2.0rc6 (#5017)
Signed-off-by: Nate Mailhot <nmailhot@nvidia.com>
|
2026-01-02 10:29:04 -08:00 |
Pavithra Vijayakrishnan
|
5034b14439
|
test: pytest mark checker (#4812)
Signed-off-by: pvijayakrish <pvijayakrish@nvidia.com>
Signed-off-by: Pavithra Vijayakrishnan <160681768+pvijayakrish@users.noreply.github.com>
|
2025-12-20 01:09:13 +00:00 |
ishandhanani
|
3e0459fb01
|
feat: bump sglang to `0.5.6.post2` and swap to upstream runtime container (#4762)
Signed-off-by: Dillon Cullinan <dcullinan@nvidia.com>
Signed-off-by: Dmitry Tokarev <dtokarev@nvidia.com>
Co-authored-by: Dillon Cullinan <dcullinan@nvidia.com>
Co-authored-by: Dmitry Tokarev <dtokarev@nvidia.com>
|
2025-12-19 18:33:04 -05:00 |
Nate Mailhot
|
910b587db1
|
chore: update nixl to 0.8.0 and ucx 1.20 (#5007)
Signed-off-by: Nate Mailhot <nmailhot@nvidia.com>
|
2025-12-19 10:55:40 -08:00 |
Anant Sharma
|
7436e3bc03
|
chore: update versions for 0.8.0 release (#5031)
Signed-off-by: Anant Sharma <anants@nvidia.com>
|
2025-12-19 10:21:52 -08:00 |
Anant Sharma
|
5b82b8b096
|
chore: update versions to match latest release (#4966)
Signed-off-by: Anant Sharma <anants@nvidia.com>
|
2025-12-18 08:45:44 -08:00 |
Jacky
|
54636097b3
|
test: Add pre-download model test that run at beginning avoiding individual test timeout (#4989)
Signed-off-by: Jacky <18255193+kthui@users.noreply.github.com>
|
2025-12-17 03:30:41 +00:00 |
Yan Ru Pei
|
d0e95c39df
|
test: remove pytest_runtestloop (#4886)
Signed-off-by: PeaBrane <yanrpei@gmail.com>
|
2025-12-11 21:29:51 +00:00 |
Guangtong Bai
|
53cec4ac22
|
feat: install Run:ai model streamer for vllm (#4848)
Signed-off-by: Guangtong Bai <guangtong.bai@gmail.com>
|
2025-12-11 20:16:43 +00:00 |
Dmitry Tokarev
|
525030324e
|
chore: TRTLLM 1.2.0rc4 (#4836)
Signed-off-by: Dmitry Tokarev <dtokarev@nvidia.com>
|
2025-12-11 03:02:24 +00:00 |
Dillon Cullinan
|
1645884d97
|
ci: fix login for operator job and other sgl/trtllm tests (#4785)
Signed-off-by: Dillon Cullinan <dcullinan@nvidia.com>
Signed-off-by: Anant Sharma <anants@nvidia.com>
Co-authored-by: Anant Sharma <anants@nvidia.com>
|
2025-12-08 13:58:27 -05:00 |
Karen Chung
|
d7c11b65e8
|
chore: bump vLLM to 0.12.0 (#4736)
Signed-off-by: alec-flowers <aflowers@nvidia.com>
Signed-off-by: Karen Chung <karenc@nvidia.com>
Signed-off-by: jthomson04 <jwillthomson19@gmail.com>
Co-authored-by: alec-flowers <aflowers@nvidia.com>
Co-authored-by: jthomson04 <jwillthomson19@gmail.com>
|
2025-12-03 19:04:43 -08:00 |
jthomson04
|
1f3e97a003
|
chore: Update sgl version (#4716)
Signed-off-by: jthomson04 <jwillthomson19@gmail.com>
Co-authored-by: ishandhanani <82981111+ishandhanani@users.noreply.github.com>
|
2025-12-03 23:35:15 +00:00 |
Karen Chung
|
60feb95549
|
chore: bump vLLM to 0.11.2 (#4476)
|
2025-12-03 14:27:27 -08:00 |
Keiven C
|
3cad926e58
|
fix: integration test failures. Support DYN_SYSTEM_PORT=0 for random port binding, and update etcd test (#4687)
Signed-off-by: Keiven Chang <keivenchang@users.noreply.github.com>
Co-authored-by: Keiven Chang <keivenchang@users.noreply.github.com>
|
2025-12-02 16:49:41 -08:00 |
Pavithra Vijayakrishnan
|
0f6dca6e70
|
test: Add pytest markers (#4111)
Signed-off-by: pvijayakrish <pvijayakrish@nvidia.com>
Signed-off-by: Pavithra Vijayakrishnan <160681768+pvijayakrish@users.noreply.github.com>
|
2025-12-01 09:53:20 -08:00 |
Tanmay Verma
|
c9d7d95f4b
|
chore: Upgrade to tensorrt-llm==1.2.0rc3 (#4645)
|
2025-11-26 22:53:41 -08:00 |
dagil-nvidia
|
ef0a4721cd
|
fix: update ai-dynamo[vllm] version and pin transformers (#4592)
Signed-off-by: Dan Gil <dagil@nvidia.com>
|
2025-11-25 21:12:16 +00:00 |
Dmitry Tokarev
|
9b8b998803
|
chore: upgrade trtllm 1.2.0rc2 (#4405)
Signed-off-by: Dmitry Tokarev <dtokarev@nvidia.com>
Co-authored-by: tanmayv25 <tanmay2592@gmail.com>
Co-authored-by: Kyle McGill <kmcgill@nvidia.com>
Co-authored-by: Ryan McCormick <rmccormick@nvidia.com>
Co-authored-by: Tanmay Verma <tanmayv@nvidia.com>
|
2025-11-21 17:27:36 -08:00 |