Optimized Heterogeneous Chiplet Composition for Hybrid LLM Workloads
Determine the optimized composition of heterogeneous compute, memory, and communication chiplets for serving hybrid Transformer–Mamba large language model workloads under architectural and runtime constraints.
References
Determining the optimized composition of heterogeneous chiplets for hybrid LLM workloads remains an open problem.
— HYDRA: A Heterogeneous Chiplet DSE Framework for Serving Dynamic Hybrid LLM Workloads
(2608.19395 - Lin et al., 19 Aug 2026) in Section 2, Related Work, page 2