WinterShiver
|
5f48dd7bf9
|
Merge 46cda3c3a9 into 713b5a3f95
|
2026-08-03 17:39:08 -07:00 |
浮梦
|
c4bbac49b2
|
[v1] support resume training from checkpoint (#10280)
docker / build (cuda) (push) Failing after 11s
Details
docker / build (npu-a2) (push) Failing after 11s
Details
docker / build (npu-a3) (push) Failing after 11s
Details
tests / tests (ubuntu-latest, 3.11, ) (push) Failing after 11s
Details
tests / tests (ubuntu-latest, 3.11, 4.55.0) (push) Failing after 11s
Details
tests / tests (ubuntu-latest, 3.11, 4.57.1) (push) Failing after 11s
Details
tests / tests (ubuntu-latest, 3.12, ) (push) Failing after 11s
Details
tests / tests (ubuntu-latest, 3.13, ) (push) Failing after 11s
Details
tests / tests (macos-latest, 3.11, ) (push) Has been cancelled
Details
tests / tests (macos-latest, 3.12, ) (push) Has been cancelled
Details
tests / tests (macos-latest, 3.13, ) (push) Has been cancelled
Details
tests / tests (windows-latest, 3.11, ) (push) Has been cancelled
Details
tests / tests (windows-latest, 3.12, ) (push) Has been cancelled
Details
tests / tests (windows-latest, 3.13, ) (push) Has been cancelled
Details
tests_cuda / tests (linux-x86_64-gpu-2, 3.11) (push) Has been cancelled
Details
tests_npu / tests (linux-aarch64-a2-4, 3.11, 2.7.1) (push) Has been cancelled
Details
Co-authored-by: frozenleaves <frozen@Mac.local>
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-04-20 20:28:08 +08:00 |
Kingsley
|
a3d44e3152
|
[mca] support qwen3.5 (#10265)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-03-10 10:55:16 +08:00 |
Philip Ottesen
|
0779846513
|
[infer] support mixed multimodal payloads (#10225)
Signed-off-by: Philip Ottesen <phiott256@gmail.com>
|
2026-02-28 20:26:53 +08:00 |
浮梦
|
f9f11dcb97
|
[v1] support training with fsdp2 (#9773)
Co-authored-by: frozenleaves <frozen@Mac.local>
Co-authored-by: Yaowei Zheng <hiyouga@buaa.edu.cn>
|
2026-01-25 19:41:58 +08:00 |
Yaowei Zheng
|
a296723697
|
[v1] upgrade batching (#9751)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-01-12 00:21:36 +08:00 |
jiaqiw09
|
81b8a50aa5
|
[deps] Update pyproject.toml and requirements (#9714)
Co-authored-by: Yaowei Zheng <hiyouga@buaa.edu.cn>
|
2026-01-04 19:52:16 +08:00 |
Yaowei Zheng
|
95ac3f2373
|
[release] Bye 2025 (#9702)
|
2025-12-31 22:22:40 +08:00 |
fivehaitao
|
c8d7e85b3e
|
[fix] Fix prediction metrics in scripts/vllm_infer.py to match Transformers (#9701)
Co-authored-by: xuht6 <xuht6@asiainfo.com>
|
2025-12-31 18:30:00 +08:00 |
Copilot
|
eceec8ab69
|
[deps] goodbye python 3.9 (#9677)
Co-authored-by: copilot-swe-agent[bot] <198982749+Copilot@users.noreply.github.com>
Co-authored-by: hiyouga <16256802+hiyouga@users.noreply.github.com>
Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>
|
2025-12-27 02:50:44 +08:00 |
WinterShiver
|
46cda3c3a9
|
[minor] rm redundant code
|
2025-11-04 10:46:31 +08:00 |
WinterShiver
|
8368f7d6d4
|
[fix] read model dtype from config
|
2025-10-31 10:21:10 +08:00 |
WinterShiver
|
b83f341276
|
[inference] formatted hf_infer script
|
2025-10-29 18:01:41 +08:00 |
WinterShiver
|
9deb38347f
|
[inference] add hf_infer script for inference using huggingface backend
|
2025-10-29 17:33:56 +08:00 |
Kingsley
|
13170577b2
|
[feat] support megatron-LM training by mcore_adapter (#9237)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
Co-authored-by: Yaowei Zheng <hiyouga@buaa.edu.cn>
|
2025-10-26 16:21:30 +08:00 |
Xiaosu Zhu
|
129e918106
|
[data] Fix Qwen3VL plugin (#9297)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
Co-authored-by: Yaowei Zheng <hiyouga@buaa.edu.cn>
Co-authored-by: kingsley <kingsleydodonow@gmail.com>
|
2025-10-26 16:07:04 +08:00 |
Yaowei Zheng
|
575e4099df
|
[misc] add qwen bench script (#9259)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2025-10-13 11:45:25 +08:00 |
Yaowei Zheng
|
6ffebe5ff7
|
[data] fix qwen omni plugin (#9204)
Co-authored-by: kingsley <kingsleydodonow@gmail.com>
|
2025-09-28 01:02:29 +08:00 |
Kingsley
|
69c9e379d5
|
[script] add Script description for qwen_omni_merge (#8293)
|
2025-06-05 13:22:01 +08:00 |
Kingsley
|
2aaede8ef4
|
[scripts] specify model class for qwen_omni merge (#8227)
|
2025-05-30 14:20:12 +08:00 |
hoshi-hiyouga
|
9b5baa97f0
|
[data] qwen3 fixes (#8109)
|
2025-05-20 02:00:30 +08:00 |
hoshi-hiyouga
|
beae231af6
|
[doc] add no build isolation (#8103)
|
2025-05-19 19:25:13 +08:00 |
Shawn Tao
|
0b773234e5
|
[infer] Modify vllm_infer.py to batch preprocess to avoid too much files opened error (#8051)
Co-authored-by: Kingsley <82590017+Kuangdd01@users.noreply.github.com>
|
2025-05-15 10:54:35 +08:00 |
Kingsley
|
9620825892
|
[scripts] add video params for vllm infer (#7992)
|
2025-05-09 21:16:52 +08:00 |
hoshi-hiyouga
|
091d2539e8
|
Merge commit from fork
|
2025-04-23 16:38:27 +08:00 |
hoshi-hiyouga
|
86ebb219d6
|
[breaking] bump transformers to 4.45.0 & improve ci (#7746)
* update ci
* fix
* fix
* fix
* fix
* fix
|
2025-04-17 02:36:48 +08:00 |
hoshi-hiyouga
|
c3c0efbaa0
|
[misc] fix packing and eval plot (#7623)
|
2025-04-07 18:20:57 +08:00 |
hoshi-hiyouga
|
831e7f1cfd
|
[model] add llama4 (#7611)
|
2025-04-06 13:42:31 +08:00 |
Kingsley
|
d32c6c014d
|
[data] fix qwen2.5 omni plugin (#7573)
* align key with qwen2vl
* nit && change scripts
|
2025-04-02 21:28:52 +08:00 |
hoshi-hiyouga
|
5e22597ff1
|
[infer] vllm video/audio inference (#7566)
|
2025-04-02 02:27:04 +08:00 |
hoshi-hiyouga
|
2bfcad2394
|
[model] fix kv cache (#7564)
|
2025-04-01 23:07:46 +08:00 |
Kingsley
|
7eed496336
|
[model] add Qwen2.5-Omni model (#7537)
* preserve image_sizes
* preserve image_sizes
* init plugin
* support audio-text2text lora
* nit
* support image/video-text2text, audio-text2text
* remove args
* remove lines
* add docs && nit
* remove some comments
* fix && add merge part script
* add license
|
2025-03-31 20:39:35 +08:00 |
hoshi-hiyouga
|
304796b803
|
[misc] fix license (#7440)
|
2025-03-23 19:31:56 +08:00 |
SnowFox4004
|
7cfd6e4bb0
|
[scripts] support compute score on vllm's predictions (#7419)
* enable manual bleu&rouge eval by adding `scripts/eval_bleu_rouge.py`
* added libraries check
* update: 使用datasets库的多进程加速处理
* update:
- 使用 fire.Fire
- 修改代码格式
* Update eval_bleu_rouge.py: correctly uses fire
Deleted the code of using sys.argv
* Update eval_bleu_rouge.py
---------
Co-authored-by: SnowFox4004 <manba@out>
Co-authored-by: hoshi-hiyouga <hiyouga@buaa.edu.cn>
|
2025-03-23 19:21:01 +08:00 |
hoshi-hiyouga
|
650a9a9057
|
[misc] update format (#7277)
|
2025-03-13 02:53:08 +08:00 |
hoshi-hiyouga
|
264538cb26
|
[misc] upgrade format to py39 (#7256)
|
2025-03-12 00:08:41 +08:00 |
hoshi-hiyouga
|
522a3e8493
|
[infer] fix vllm args (#7235)
Former-commit-id: 999be5b4512890b8cf4f45874a77e35cf35626f5
|
2025-03-11 01:15:35 +08:00 |
hoshi-hiyouga
|
f4ec4fa6ad
|
[script] fix vllm version (#7193)
Former-commit-id: ababdde597b2b9bf0ab3f30f036bc8d97de07f03
|
2025-03-06 17:14:17 +08:00 |
hoshi-hiyouga
|
5f65558088
|
[misc] fix project toml (#7067)
Former-commit-id: 28a668ff4e0beebfe5387362f5518c1d9343666f
|
2025-02-25 23:22:48 +08:00 |
JieShen
|
0f54a78144
|
[script] add seed args (#7058)
* add seed args
* add seed args
* update seed
Former-commit-id: eb9770b2c01a840b6a0ac119210c22bdbb81e18b
|
2025-02-25 19:44:57 +08:00 |
hoshi-hiyouga
|
be33ef67fb
|
[misc] fix script (#6977)
Former-commit-id: 775efa1d8cbdb1b7d122be2a986d47f85214e0a1
|
2025-02-18 17:00:46 +08:00 |
hoshi-hiyouga
|
3a3f4072e5
|
[misc] fix grad ckpt func (#6916)
Former-commit-id: 35e069a52b3d7cfd9b0107574b09265eb2290f0b
|
2025-02-13 00:17:18 +08:00 |
hoshi-hiyouga
|
88eafd865b
|
[misc] support export ollama modelfile (#6899)
* support export ollama modelfile
* update config
* add system and num ctx
Former-commit-id: 8c2af7466f4015f300b51841db11bcd2505ebf20
|
2025-02-11 19:52:25 +08:00 |
codingma
|
b72c4bd118
|
support ollama modelfile export (#4686)
Former-commit-id: 15cca102a7fc0d08b5d049cf264acc6fa576b104
|
2025-02-11 17:52:24 +08:00 |
Zhangchi Feng
|
8f401e37f8
|
[model] support audio (#6701)
* support qwen2_audio
* improve code
* lint
* fix
* fix
* fix
---------
Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>
Former-commit-id: 5eacb5629e4d7733cd992a63747a1335f2c6a929
|
2025-02-05 04:59:09 +08:00 |
hoshi-hiyouga
|
c2022431aa
|
[misc] update license year & fix llama pro (#6814)
* fix llamapro script
* change year
Former-commit-id: d9ae594178796994d400a5f207d6499712816f89
|
2025-02-05 01:53:33 +08:00 |
hoshi-hiyouga
|
28d145a066
|
pin vllm version to 0.6.5 (#6629)
Former-commit-id: 26097ca0adf25ebb7d9e8eec2d2cef673c6cfe88
|
2025-01-14 02:44:02 +08:00 |
hoshi-hiyouga
|
2a05941b14
|
[inference] fix stop token for object detection (#6624)
* fix stop token
* update minicpm data pipeline
* fix npu qlora examples
Former-commit-id: 844919fadaa8a61dfae47020971ea80730b2346f
|
2025-01-13 21:34:20 +08:00 |
hiyouga
|
8516054e4d
|
update scripts
Former-commit-id: 05aa52adde8905ca892f1ed5847d6f90b1992848
|
2025-01-03 10:50:32 +00:00 |
hiyouga
|
bbd432415d
|
support qwen2vl vllm infer
Former-commit-id: 03ddd2555fb97488cd4daab11e8b672d36150c5a
|
2024-12-05 10:17:26 +00:00 |