基于PPO的逻辑综合序列优化通用框架
PDF下载 (161)王梦可,杨朝晖,查晓婧,夏银水*.基于PPO的逻辑综合序列优化通用框架[J].宁波大学学报(理工版),2025,38(2):78-85.DOI:10.20098/j.cnki.1001-5132.2024.0204
WANG Mengke,YANG Zhaohui,ZHA Xiaojing,XIA Yinshui.PPO-based universal framework design for logic synthesis sequence optimization[J].Journal of Ningbo University(Natural Science & Engineering Edition),2025,38(2):78-85.DOI:10.20098/j.cnki.1001-5132.2024.0204
| Title: | PPO-based universal framework design for logic synthesis sequence optimization |
| 作者: | 王梦可, 杨朝晖, 查晓婧, 夏银水* |
| Author(s): | WANG Mengke, YANG Zhaohui, ZHA Xiaojing, XIA Yinshui |
| 关键词: | 逻辑综合; 序列优化; 强化学习; 近端策略优化 |
| Keywords: | logic synthesis; sequence optimization; reinforcement learning; proximal policy optimization |
| 分类号: | TP331.2; TN431.2 |
| DOI: | 10.20098/j.cnki.1001-5132.2024.0204 |
| 文献标识码: | A |
| 摘要: | 逻辑综合通常采用启发式方法将逻辑优化算法组成为序列进行电路性能优化, 而启发式方法难以根据电路和优化目标的差异进行序列自动化调节, 影响了电路优化质量. 为了在集成电路设计中提升序列的自适应生成能力, 将序列优化问题建模为马尔可夫决策过程, 提出一种面向多种逻辑表示的强化学习框架, 利用近端策略优化(Proximal Policy Optimization, PPO)指导智能体来探索序列优化空间, 改善其生成序列的泛化能力. 并将EPFL基准电路转变为与-非图(And-Inverter Graph, AIG)和异或多数图(Xor-Majority Graph, XMG)形式, 分别经由所提出的框架进行实验, AIG形式下本文方法与DRiLLS和BOiLS方法相比分别有18.66百分点和27.67百分点的性能提升; XMG形式下则可提升原始电路性能约37.34%. 实验结果表明, 由本文方法生成的算法序列对电路性能有较大改进. |
| Abstract: | Logic synthesis typically employs heuristic methods to compose a sequence of logic optimization algorithms for circuit performance improvement. However, heuristic methods face challenges in automatically adjusting sequences based on circuit and optimization objectives, thus affecting the quality of circuit optimization. In order to enhance the capability of adaptive sequence generation in integrated circuit design, this paper models the sequence optimization problem as a Markov Decision Process and proposes a reinforcement learning framework for multiple logic representations. The framework utilizes Proximal Policy Optimization (PPO) to guide the agent in exploring the sequence optimization space and improve the generalization ability of its sequence generation. The paper transforms EPFL benchmark circuits into AIG and XMG forms, respectively, and conducts experiments on them by using the proposed framework. Compared to DRiLLS and BOiLS methods, the proposed method achieves performance improvements of 18.66 percentages and 27.67 percentages in the AIG form, and approximately 37.34% improvement in the original circuit performance in the XMG form. The experimental results demonstrate significant improvements in circuit performance achieved by the algorithm sequences generated with the proposed method |
| 参考文献 /References: | [1] 储著飞, 王伦耀, 夏银水. 基于多逻辑域的逻辑综合研究进展[J]. 微纳电子与智能制造, 2021, 3(2):64-73.
[2] LI X, CHEN L, YANG F, et al. HIMap: a heuristic and iterative logic synthesis approach[C]//Proceedings of the 59th ACM/IEEE Design Automation Conference. San Francisco, CA, USA. ACM, 2022:415-420. [3] CHEN H Z, SHEN M H. A deep-reinforcement-learning- based scheduler for high-level synthesis[C]//Proceedings of the 2019 ACM/SIGDA International Symposium on Field-Programmable Gate Arrays. Seaside, CA, USA. ACM, 2019:117. [4] NETO W L, LI Y, GAILLARDON P E, et al. Flowtune: end-to-end automatic logic optimization exploration via domain-specific multi-armed bandit[J]. IEEE Transac- tions on Computer-Aided Design of Integrated Circuits and Systems, 2022, 42(6):1912-1925. [5] AGNESINA A, CHANG K, LIM S K. VLSI placement parameter optimization using deep reinforcement learning [C]//2020 IEEE/ACM International Conference on Computer Aided Design (ICCAD). San Diego, CA, USA. IEEE, 2020:1-9. [6] YU C X, XIAO H P, DE MICHELI G. Developing synthesis flows without human knowledge[C]//2018 55th ACM/ESDA/IEEE Design Automation Conference (DAC). San Francisco, CA, USA. IEEE, 2018:1-6. [7] YANG C H, XIA Y S, CHU Z F, et al. Logic synthesis optimization sequence tuning using RL-based LSTM and graph isomorphism network[J]. IEEE Transactions on Circuits and Systems II: Express Briefs, 2021, 69(8): 3600-3604. [8] GU Y, CHENG Y H, CHEN C L P, et al. Proximal policy optimization with policy feedback[J]. IEEE Transactions on Systems, Man, and Cybernetics: Systems, 2022, 52(7): 4600-4610. [9] BRAYTON R, MISHCHENKO A. ABC: an academic industrial-strength verification tool[C]//International Conference on Computer Aided Verification. Berlin, Heidelberg: Springer, 2010:24-40. [10] CHU Z. Advanced logic synthesis and optimization tool (ALSO) (Release 20190221)[EB/OL]. [2024-02-18]. https://gitee.com/zfchu/also. [11] MISHCHENKO A, CHATTERJEE S, BRAYTON R. DAG-aware AIG rewriting a fresh look at combinational logic synthesis[C]//2006 43rd ACM/IEEE Design Automation Conference. San Francisco, CA, USA. IEEE, 2006:532-535. [12] AMARÚ L, GAILLARDON P E, DE MICHELI G. Majority-inverter graph: a novel data-structure and algorithms for efficient logic optimization[C]//2014 51st ACM/EDAC/IEEE Design Automation Conference (DAC). San Francisco, CA, USA. IEEE, 2014:1-6. [13] CHU Z F, SOEKEN M, XIA Y S, et al. Structural rewriting in XOR-majority graphs[C]//Proceedings of the 24th Asia and South Pacific Design Automation Conference. Tokyo, Japan. ACM, 2019:663-668. [14] SHAH D, HUNG E, WOLF C, et al. Yosys+ nextpnr: an open source framework from verilog to bitstream for commercial FPGAs[C]//2019 IEEE 27th Annual International Symposium on Field-Programmable Custom Computing Machines (FCCM). San Diego, CA, USA. IEEE, 2019:1-4. [15] AMARÚ L, POSSANI V, TESTA E, et al. LUT-based optimization for ASIC design flow[C]//2021 58th ACM/IEEE Design Automation Conference (DAC). San Francisco, CA, USA. IEEE, 2021:871-876. [16] PASZKE A, GROSS S, MASSA F, et al. PyTorch: an imperative style, high-performance deep learning library [EB/OL]. [2024-01-12]. http://arxiv.org/abs/1912.01703. [17] AMARÚ L, GAILLARDON P E, DE MICHELI G. The EPFL combinational benchmark suite[C]//Proceedings of the 24th International Workshop on Logic & Synthesis (IWLS). 2015:1-5. [18] HOSNY A, HASHEMI S, SHALAN M, et al. DRiLLS: deep reinforcement learning for logic synthesis[C]//2020 25th Asia and South Pacific Design Automation Conference (ASP-DAC). Beijin, China, IEEE, 2020:581- 586. [19] GROSNIT A, MALHERBE C, TUTUNOV R, et al. BOiLS: Bayesian optimisation for logic synthesis [C]//2022 Design, Automation & Test in Europe Conference & Exhibition (DATE). Antwerp, Belgium. IEEE, 2022:1193-1196. [20] MISHCHENKO A, CHATTERJEE S, BRAYTON R. DAG-aware AIG rewriting: a fresh look at combinational logic synthesis[C]//2006 43rd ACM/IEEE Design Automation Conference. San Francisco, CA, USA. IEEE, 2006:532-535. |
| 备注/Memo: | 收稿日期: 2024−02−18. 宁波大学学报(理工版)网址: http://journallg.nbu.edu.cn/ 基金项目: 国家自然科学基金(62131010, U22A2013); 国家自然科学基金青年项目(62304115); 浙江省自然科学基金创新群体课题(LDT23F04021F04); 浙江省科研计划一般项目(Y202248965). 第一作者: 王梦可, 硕士研究生, 主要研究方向: 集成电路电子设计自动化. E-mail: 2111082020@nbu.edu.cn *通信作者: 夏银水, 教授, 主要研究方向: 低功耗集成电路设计及电子设计自动化. E-mail: xiayinshui@nbu.edu.cn 宁波大学学报(理工版)网址:http://journallg.nbu.edu.cn/ |