sglang

mirror of https://github.com/kvcache-ai/sglang.git synced 2026-07-01 04:08:10 +00:00

Author	SHA1	Message	Date
Wenyao Gao	4dfc8e1c3f	VLM: support passing --mm-process-config for all models (#18467 )	2026-04-12 17:08:05 +08:00
Aditya Sharma	f6e85676b5	model: support qwen3-asr (#22073 ) Co-authored-by: Xinyuan Tong <115166877+JustinTong0323@users.noreply.github.com>	2026-04-07 13:27:05 +08:00
Piotr Mazurek	b5e8c4b9e3	model: support LFM2-VL (Liquid Foundation Model 2 Vision-Language) (#21230 ) Co-authored-by: Piotr Mazurek <piotr.mazurek@liquid.ai>	2026-04-04 16:36:04 +08:00
Nave Assaf	77872a8d55	Update Nemotron Example docs to include Super v3 and Nano 4B (#21416 ) Signed-off-by: Nave Assaf <nassaf@nvidia.com>	2026-03-25 12:03:19 -04:00
Jiaxin(Jackson) Deng	c4db64c16b	Add Lychee Doc Links Check to Local and CI (#19742 ) Co-authored-by: Zijie Xia <zijie_xia@icloud.com> Co-authored-by: Zijie Xia <zijiexia@users.noreply.github.com> Co-authored-by: zijiexia <37504505+zijiexia@users.noreply.github.com>	2026-03-24 13:48:26 -07:00
rakesh	a710b7d791	[Sarvam] Add inference support for Sarvam MoE LLMs (#18938 ) Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>	2026-03-04 15:28:00 -08:00
Michael	6b8e62f94f	[AMD] [Qwen 3.5 Day 0] Add Qwen 3.5 nightly accuracy tests (#19479 )	2026-03-02 19:42:42 -08:00
Michael	403195d59d	[AMD] [MiniMax-M2.5 Day 0] Add MiniMax-M2.5 nightly accuracy test (#19443 )	2026-02-27 02:39:33 -08:00
chengshuang18	295bc17576	Feature/sdar support (#19044 ) Co-authored-by: root <root@gpu-lg-cmc-h-h200-3047.host.h.pjlab.org.cn> Co-authored-by: chengshuang <chengshuang@pjlab.org.cn> Co-authored-by: 赵晨阳 <zhaochen20@outlook.com>	2026-02-19 21:58:15 -08:00
Cheng Wan	73a7f0d049	Revert "Add SDAR model support" (#19032 )	2026-02-19 16:03:56 -08:00
chengshuang18	44ab752b7a	Add SDAR model support (#18318 ) Co-authored-by: root <root@gpu-lg-cmc-h-h200-3047.host.h.pjlab.org.cn> Co-authored-by: chengshuang <chengshuang@pjlab.org.cn> Co-authored-by: 赵晨阳 <zhaochen20@outlook.com>	2026-02-19 11:20:32 -08:00
Bhavneek Singh	1ce3420784	Model: Support IBM Granite (Dense/Mamba + MoE) (#18040 )	2026-02-15 11:24:41 +08:00
qianyue76	f06ab17a73	[diffusion] docs: consolidate diffusion documentation into docs (#18095 ) Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com> Co-authored-by: JiaxinD <djx2048@gmail.com>	2026-02-11 16:55:07 -08:00
brimon	ddbcfbaaab	feature: support bidirectional attention for Gemma-3 (#10707 )	2026-02-09 23:17:45 +08:00
Junlin Zhou	14652243bd	[DLLM] Add JointThreshold algorithm for joint M2T and T2T decoding (#18171 ) Signed-off-by: Junlin Zhou <zhoujunlin.zjl@antgroup.com> Co-authored-by: Tiwei Bie <tiwei.btw@antgroup.com>	2026-02-09 14:20:45 +08:00
Rishit Shivam	c850a8a41a	[Docs] Add Falcon H1, Hunyuan-Large, Qwen3-Omni support and update Diffusion usage (#17888 ) Co-authored-by: Rishitshivam <164783543+Rishitshivam@users.noreply.github.com> Co-authored-by: Ratish P <114130421+Ratish1@users.noreply.github.com> Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com> Co-authored-by: Adarsh Shirawalmath <114558126+adarshxs@users.noreply.github.com> Co-authored-by: zhaochenyang20 <zhaochen20@outlook.com>	2026-02-06 13:17:51 -08:00

16 Commits