DICE: Detailed Inter-Chiplet End-to-End PHY Modeling for Accurate Chiplet Simulation
R. Aligholipour, S. Kaxiras and Yuan Yao, "DICE: Detailed Inter-Chiplet End-to-End PHY Modeling for Accurate Chiplet Simulation," arXiv:2607.24221, 2026.
Assistant Professor, Computer Architecture · Uppsala University · yuan.yao@it.uu.se
R. Aligholipour, S. Kaxiras and Yuan Yao, "DICE: Detailed Inter-Chiplet End-to-End PHY Modeling for Accurate Chiplet Simulation," arXiv:2607.24221, 2026.
R. Aligholipour, S. Kaxiras and Yuan Yao, "DICE: Detailed Inter-Chiplet End-to-End PHY Modeling for Accurate Chiplet Simulation," 2026 ACM/IEEE 53rd Annual International Symposium on Computer Architecture (ISCA), Raleigh, USA, 2026.
Yuan Yao, S. Li, R. Aligholipour and S. Kaxiras, "NoCWalk: In-Network Page Walks for Concurrent Data Structure Workloads on Multicore," 2026 ACM International Conference on Computing Frontiers (CF), Catania, Italy, 2026.
S. Li, L. Mottola, Yuan Yao and S. Kaxiras, "ConvReflex: Efficient Ultra-Low-Power CNN Inference via Clamping Prediction," ACM/IEEE International Conference on Embedded Artificial Intelligence and Sensing Systems (SenSys), Saint-Malo, France, 2026.
J. Söderström, R. Aligholipour and Yuan Yao, "Understanding Simulated Architecture via gem5 Call-Stack Profiling," 2026 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS), Seoul, South Korea, 2026.
H. Atmer, Yuan Yao, T. Voigt, and S. Kaxiras, "Prefill vs. Decode Bottlenecks: SRAM-Frequency Tradeoffs and the Memory-Bandwidth Ceiling," arXiv:2512.22066, 2025.
R. Aligholipour and Yuan Yao, "The Fake-Busy and True-Idle Problems of Running Graph Applications on Chiplet-Based Multi-Cores," 2025 IEEE International Symposium on Workload Characterization (IISWC), Irvine, USA, 2025.
J. Söderström and Yuan Yao, "Anatomy of the gem5 Simulator: AtomicSimpleCPU, TimingSimpleCPU, O3CPU, and Their Interaction with the Ruby Memory System," arXiv:2508.18043, 2025.
R. Aligholipour, P. Aimoniotis, S. Kaxiras and Yuan Yao, "RXT: RefleXive address Translation for Pointer-Chasing Workloads," 39th IEEE International Parallel & Distributed Processing Symposium (IPDPS), Milan, Italy, 2025.
Yuan Yao, X. Chen, H. Atmer and S. Kaxiras, "TangramFP: Energy-Efficient, Bit-Parallel, Multiply-Accumulate for Deep Neural Networks," 2024 IEEE 36th International Symposium on Computer Architecture and High Performance Computing (SBAC-PAD), Hilo, HI, USA, 2024, pp. 1-12, doi: 10.1109/SBAC-PAD63648.2024.00009.
Weining Song, Stefanos Kaxiras, Thiemo Voigt, Yuan Yao, and Luca Mottola. 2024. TaDA: Task Decoupling Architecture for the Battery-less Internet of Things. In Proceedings of the 22nd ACM Conference on Embedded Networked Sensor Systems (SenSys). Association for Computing Machinery, New York, NY, USA, 409–421. https://doi.org/10.1145/3666025.3699347
Weining Song, Stefanos Kaxiras, Luca Mottola, Thiemo Voigt, and Yuan Yao, ''Silent Stores in the Battery-less Internet of Things: A Good Idea?'' in Proceedings of the 2023 International Conference on embedded Wireless Systems and Networks (EWSN), Association for Computing Machinery, New York, NY, USA, 40-45.
Yuan Yao, "SE-CNN: Convolution Neural Network Acceleration via Symbolic Value Prediction," in IEEE Journal on Emerging and Selected Topics in Circuits and Systems, vol. 13, no. 1, pp. 73-85, March 2023, doi: 10.1109/JETCAS.2023.3244767.
Yuan Yao, "Game-of-Life Temperature-Aware DVFS Strategy for Tile-Based Chip Many-Core Processors," in IEEE Journal on Emerging and Selected Topics in Circuits and Systems, vol. 13, no. 1, pp. 58-72, March 2023, doi: 10.1109/JETCAS.2023.3244763
P. Ekemark, Yuan Yao, A. Ros, K. Sagonas and S. Kaxiras, "TSOPER: Efficient Coherence-Based Strict Persistency," 2021 IEEE International Symposium on High-Performance Computer Architecture (HPCA), Seoul, Korea (South), 2021, pp. 125-138, doi: 10.1109/HPCA51647.2021.00021.
Yuan Yao and Z. Lu, "Pursuing Extreme Power Efficiency With PPCC Guided NoC DVFS," in IEEE Transactions on Computers, vol. 69, no. 3, pp. 410-426, 1 March 2020, doi: 10.1109/TC.2019.2949807.
Z. Lu and Yuan Yao, "Thread Voting DVFS for Manycore NoCs," in IEEE Transactions on Computers, vol. 67, no. 10, pp. 1506-1524, 1 Oct. 2018, doi: 10.1109/TC.2018.2827039.
Yuan Yao and Z. Lu, "iNPG: Accelerating Critical Section Access with In-network Packet Generation for NoC Based Many-Cores," 2018 IEEE International Symposium on High Performance Computer Architecture (HPCA), Vienna, 2018, pp. 15-26, doi: 10.1109/HPCA.2018.00012.
Z. Lu and Yuan Yao, "Marginal Performance: Formalizing and Quantifying Power Over/Under Provisioning in NoC DVFS," in IEEE Transactions on Computers, vol. 66, no. 11, pp. 1903-1917, 1 Nov. 2017, doi: 10.1109/TC.2017.2715018.
Z. Lu and Yuan Yao, "Dynamic Traffic Regulation in NoC-Based Systems," in IEEE Transactions on Very Large Scale Integration (VLSI) Systems, vol. 25, no. 2, pp. 556-569, Feb. 2017, doi: 10.1109/
Zhonghai Lu and Yuan Yao. 2016. Aggregate Flow-Based Performance Fairness in CMPs. ACM Trans. Archit. Code Optim. 13, 4, Article 53 (December 2016), 27 pages. https://doi.org/10.1145/3014429
Yuan Yao and Z. Lu, "Opportunistic Competition Overhead Reduction for Expediting Critical Section in NoC Based CMPs," 2016 ACM/IEEE 43rd Annual International Symposium on Computer Architecture (ISCA), Seoul, Korea (South), 2016, pp. 279-290, doi: 10.1109/ISCA.2016.33.
Yuan Yao and Z. Lu, "Memory-access aware DVFS for network-on-chip in CMPs," 2016 Design, Automation & Test in Europe Conference & Exhibition (DATE), Dresden, Germany, 2016, pp. 1433-1436.
Yuan Yao and Z. Lu, "DVFS for NoCs in CMPs: A thread voting approach," 2016 IEEE International Symposium on High Performance Computer Architecture (HPCA), Barcelona, Spain, 2016, pp. 309-320, doi: 10.1109/HPCA.2016.7446074.
Z. Lu, Yuan Yao and Y. Jiang, "Towards stochastic delay bound analysis for Network-on-Chip," 2014 Eighth IEEE/ACM International Symposium on Networks-on-Chip (NoCS), Ferrara, Italy, 2014, pp. 64-71, doi: 10.1109/NOCS.2014.7008763.
Yuan Yao and Z. Lu, "Fuzzy Flow Regulation for Network-on-Chip based Chip Multiprocessors systems," 2014 19th Asia and South Pacific Design Automation Conference (ASP-DAC), Singapore, 2014, pp. 343-348, doi: 10.1109/ASPDAC.2014.6742913.
Conference presentation at ISCA 2026, Raleigh, North Carolina, USA
Invited seminar at IEEE CEDA Guangzhou Chapter (HKUST), Guangzhou, China
Workshop talk at gem5 Workshop at ISCA 2025, Tokyo, Japan
Workshop talk at gem5 Workshop at ISCA 2025, Tokyo, Japan