Logo
Explore Help
Sign In
apps/FastDeploy
1
0
Fork 0
You've already forked FastDeploy
mirror of https://github.com/PaddlePaddle/FastDeploy.git synced 2026-04-23 17:11:21 +08:00
Code Issues Actions 10 Packages Projects Releases Wiki Activity
Files
fb2eb403abfa4a8db7c8d8ac4231d50d6561a4c9
FastDeploy/fastdeploy/model_executor
T
History
Sunny-bot1 59d2edde29 [BugFix] Add support for weight shape constraints and group size selection in Machete (#4911)
2025-11-10 20:57:35 +08:00
..
graph_optimization
[Graph Optimization] Refactor default capture list (#4617)
2025-10-28 21:31:02 +08:00
guided_decoding
[FDConfig]Remove reasoning_parser/guided_decoding_backend/disable_any_whitespace/device_ids in FDConfig (#4362)
2025-10-17 10:40:59 +08:00
layers
[BugFix] Add support for weight shape constraints and group size selection in Machete (#4911)
2025-11-10 20:57:35 +08:00
logits_processor
[Feature] support logits processors (#4515)
2025-10-29 00:08:53 +08:00
model_loader
[Speculative Decoding][MTP]Support mtp in epdptp mode (#4614)
2025-10-28 16:02:47 +08:00
models
[Metax] support ERNIE-4.5-VL-28B (#4820)
2025-11-07 04:55:49 -08:00
ops
delete useless code (#4544)
2025-10-23 13:40:34 +08:00
__init__.py
…
forward_meta.py
remove input_ids from ForwardMeta (#4793)
2025-11-05 11:55:51 +08:00
load_weight_utils.py
[XPU] ep+tp all2all (#4836)
2025-11-06 17:26:14 +08:00
pre_and_post_process.py
[Metax] support ERNIE-4.5-VL-28B (#4820)
2025-11-07 04:55:49 -08:00
utils.py
[PD Disaggregation] Support Qwen3-MoE use PD + EP inference. (#4691)
2025-11-06 10:32:15 +08:00
Powered by Gitea Version: 1.26.0 Page: 1223ms Template: 6ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API