TensorRT-LLMs/docs/source/blogs/media
Fanrong Li 862bde99b6
draft[doc]: add mtp tech blog (#4580)
* add mtp tech blog.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* update figure size.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* update the figure caption style and add some code/pr links.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix figure captions.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix figure size and perf data.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

* fix based on comments

Signed-off-by: Yue Weng <25103990+yweng0828@users.noreply.github.com>

* fix figure links.

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>

---------

Signed-off-by: Fanrong Li <23290157+lfr-0531@users.noreply.github.com>
Signed-off-by: Yue Weng <25103990+yweng0828@users.noreply.github.com>
Co-authored-by: Yue Weng <25103990+yweng0828@users.noreply.github.com>
2025-05-23 13:54:21 +08:00
..
.gitkeep Add Latest News section (#315) 2023-11-08 15:04:33 +08:00
Falcon180B-H200_acc.png Update latest news (#549) 2023-12-04 22:04:00 +08:00
Falcon180B-H200_DecvOct.png Update latest news (#549) 2023-12-04 22:04:00 +08:00
Falcon180B-H200_H200vA100.png Update latest news (#549) 2023-12-04 22:04:00 +08:00
Falcon180B-H200_tps.png Update latest news (#549) 2023-12-04 22:04:00 +08:00
H200launch_H200vsH100_tps.png Add Latest News section (#362) 2023-11-13 15:17:23 +08:00
H200launch_tps.png Add Latest News section (#365) 2023-11-13 20:56:22 +08:00
moe_structure.png Update TensorRT-LLM (#1358) 2024-03-26 20:47:14 +08:00
tech_blog1_fuse_a_gemm.png doc: DS r1 min latency blog (#4386) 2025-05-16 20:20:28 +08:00
tech_blog1_model_details.png doc: DS r1 min latency blog (#4386) 2025-05-16 20:20:28 +08:00
tech_blog1_model_overview.png doc: DS r1 min latency blog (#4386) 2025-05-16 20:20:28 +08:00
tech_blog1_router_gemm.png doc: DS r1 min latency blog (#4386) 2025-05-16 20:20:28 +08:00
tech_blog1_sparse_exp_as_a_gemm.png doc: DS r1 min latency blog (#4386) 2025-05-16 20:20:28 +08:00
tech_blog2_acc_relaxed_acceptance.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_mtp_eagle.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_mtp_modules.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_mtp_vanilla.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_overall_workflow.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_perf_and_ar.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_relaxed_acceptance.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_tree_spec_decoding.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tech_blog2_verify_and_accept.png draft[doc]: add mtp tech blog (#4580) 2025-05-23 13:54:21 +08:00
tp_ep.png Update TensorRT-LLM (#1358) 2024-03-26 20:47:14 +08:00
TRT_LLM_v0-5-0_H100vA100_1st.png Add Latest News section (#315) 2023-11-08 15:04:33 +08:00
TRT_LLM_v0-5-0_H100vA100_tps.png Add Latest News section (#315) 2023-11-08 15:04:33 +08:00
XQA_ThroughputvsLatency.png Doc update 20240130 (#1009) 2024-01-31 03:40:22 +08:00