verilator

Commit Graph

Author	SHA1	Message	Date
Geza Lore	6bc48fcdb3	Improve Dfg type system (#6390 ) Added a mini type system for Dfg using DfgDataType to replace Dfg's use of AstNodeDType. This is much more restricted and represents only the types Dfg can handle in a canonical form. This will be needed when adding more support for unpacked arrays and maybe unpacked structs one day. Also added an internal type checker for DfgGraphs which encodes all the assumptions the code makes about type relationships in the graph. Run this in a few places with --debug-check. Fix resulting fallout.	2025-09-07 20:38:50 +01:00
Geza Lore	a6f26b85b3	Internals: Improve DFG implementation details (#6355 ) Large scale refactoring to simplify some of the more obtuse internals of DFG. Remove multiple redundant internal APIs, simplify representation of variables, fix potential unsoundness in circular decomposition. No functional change intended.	2025-09-02 16:50:40 +01:00
Geza Lore	3770273637	Optimize logic in non-virtual interfaces with DFG (#6347 )	2025-08-30 10:35:16 -04:00
Geza Lore	02e64f0795	Optimize multiplexers in Dfg synthesis (#6331 ) The previous algorithm was designed to handle the general case where a full control flow path predicate is required to select which value to use when synthesizing control flow join point in an always block. Here we add a better algorithm that tries to use the predicate of the closest dominating branch if the branch paths dominate the joining paths. This is almost universally true in synthesizable logic (RTLMeter has no exceptions), however there are cases where this is not applicable, for which we fall back on the previous generic algorithm. Overall this significantly simplifies the synthesized Dfg graphs and enables further optimization.	2025-08-25 13:47:45 +01:00
Geza Lore	636a6b8cd2	Optimize complex combinational logic in DFG (#6298 ) This patch adds DfgLogic, which is a vertex that represents a whole, arbitrarily complex combinational AstAlways or AstAssignW in the DfgGraph. Implementing this requires computing the variables live at entry to the AstAlways (variables read by the block), so there is a new ControlFlowGraph data structure and a classical data-flow analysis based live variable analysis to do that at the variable level (as opposed to bit/element level). The actual CFG construction and live variable analysis is best effort, and might fail for currently unhandled constructs or data types. This can be extended later. V3DfgAstToDfg is changed to convert the Ast into an initial DfgGraph containing only DfgLogic, DfgVertexSplice and DfgVertexVar vertices. The DfgLogic are then subsequently synthesized into primitive operations by the new V3DfgSynthesize pass, which is a combination of the old V3DfgAstToDfg conversion and new code to handle AstAlways blocks with complex flow control. V3DfgSynthesize by default will synthesize roughly the same constructs as V3DfgAstToDfg used to handle before, plus any logic that is part of a combinational cycle within the DfgGraph. This enables breaking up these cycles, for which there are extensions to V3DfgBreakCycles in this patch as well. V3DfgSynthesize will then delete all non synthesized or non synthesizable DfgLogic vertices and the rest of the Dfg pipeline is identical, with minor changes to adjust for the changed representation. Because with this change we can now eliminate many more UNOPTFLAT, DFG has been disabled in all the tests that specifically target testing the scheduling and reporting of circular combinational logic.	2025-08-19 15:06:38 +01:00
Wilson Snyder	88046c8063	Internals: Rename AstSenTree pointers to sentreep. No functional change intended except JSON.	2025-08-17 19:14:34 -04:00
Geza Lore	d273e2cbd0	Internals: Do not astgen useless Dfg vertex subtypes	2025-08-15 20:06:58 +01:00
Geza Lore	16d32cdd4a	Internals: Refactor Ast to Dfg conversion for reusability. (#6276 ) This is mainly code motion, with minimal algorithmic changes to facilitate reusing parts in future code. No functional change intended.	2025-08-08 22:53:12 +01:00
Geza Lore	d2edab458e	Refactor Dfg variable flags (#6259 ) Store all flags in a DfgVertexVar relating to the underlying AstVar/AstVarScope stored via AstNode::user1(). user2/user3/user4 are then usable by DFG algorithms as needed.	2025-08-05 10:24:54 +01:00
Geza Lore	deed20fb78	Fix partial DFG conversion of concat assignments (#6255 ) When we had a `{a, b} = ...`, and the DFG conversion of `a = ...` succeeded, but `b = ...` failed, we still used to include `a = ...` in the DFG, which then caused a spurious multi-driver error for `a` on a subsequent DFG pass, as the original `{a, b} = ...` was still present in the Ast, but we also had the extra `a = ...` from converting out of DFG on the previous pass. In this patch we only convert assignments with a concatenation on the LHS, if all target LValues can be converted into DFG. This is the proper fix for #4231	2025-08-03 14:52:20 +01:00
Geza Lore	504884b7d5	Refactor DFG context objects (#6232 ) - Move All DFG context objects to V3DfgContext.h - Add separate object for ast2dfg and dfg2ast passes - Factor out commonalities No functional change	2025-07-26 20:37:01 +01:00
Geza Lore	7646e7d89c	Exclude SystemC variables from DFG (#6208 ) SystemC variables are fairly special (they can only be assigned to/from, but not otherwise participate in expressions), which complicates some DFG code. These variables only ever appear as port on the top level wrapper, so excluding them from DFG does not make us loose any optimizations, but simplifies internals.	2025-07-21 18:32:08 +01:00
Geza Lore	a8dca71ed0	Support more complex combinational assignments in DFG. (#6205 ) Previously DFG was limited to having a Sel, or an ArraySel potentially under a Concat on the LHS of combinational assignments. Other forms or combinations were not representable in the graph. This adds support for arbitrary combinations of the above by combining DfgSplicePacked and DfgSpliceArray vertices introduced in #6176. In particular, Sel(ArraySel(VarRef,_),_) enables a lot more code to be represented in DFG.	2025-07-21 12:33:12 -04:00
Geza Lore	03e0d49d99	Optimize DFG partial assignments (#6176 ) This is mostly a refactoring, but also enables handling some more UNOPTFLAT, when the variable is only partially assigned in the cycle. Previously the way partial assignments to variables were handled were through the DfgVerexVar types themselves, which kept track of all drivers. This has been replaced by DfgVertexSplice (which always drives a DfgVeretexVar), and all DfgVertexVar now only have a single source, either a DfgVertexSplice, if partially assigned, or an arbitrary DfgVertex when updated as a whole.	2025-07-14 17:09:34 -04:00
Geza Lore	7a3f1f16ca	Optimize DFG before V3Gate (#6141 )	2025-07-01 17:55:08 -04:00
Wilson Snyder	b914cda1c7	Internals: cppcheck cleanups. No functional change.	2025-06-28 12:29:41 -04:00
Geza Lore	916d473eff	Internals: Replace unnecessary AstSel::widthp() child node with const in node (#6117 )	2025-06-24 11:59:09 -04:00
Wilson Snyder	8fbb725f34	Copyright year update.	2025-01-01 08:30:25 -05:00
Wilson Snyder	0c820c3068	Internals: Standardize template argument names. No functional change.	2024-11-29 20:20:38 -05:00
Geza Lore	cf111d2e1f	Do not create aliases for forced port signals (#5105 ) + don't remove forced signals in V3Const and Dfg Fixes #5062	2024-05-10 18:19:51 +01:00
Geza Lore	745605efe3	Fix DFG removing forceable signals (#4942 ) DFG could remove forceable signals by replacing them with their in-design driver. This is a bit of a pain to prevent, and ideally the forcing transform should happen before DFG, but implementing it there is a pain due to having to rewrite ports based on direction. This is an attempted fix in DFG. More cases might remain.	2024-03-03 16:22:41 +00:00
Wilson Snyder	3a5248a919	Internals: Mark structs final/VL_NOT_FINAL. No functional change intended.	2024-01-20 15:06:46 -05:00
Wilson Snyder	e76f29e5ba	Copyright year update	2024-01-01 03:19:59 -05:00
Wilson Snyder	f3ae4b8786	Fix spelling	2023-11-10 23:25:53 -05:00
Wilson Snyder	b5828a7ce9	Fix header order botched by clang-format in recent commit.	2023-10-18 06:37:46 -04:00
github action	770cd24f27	Apply 'make format'	2023-10-18 02:50:27 +00:00
Wilson Snyder	431bb1ed16	Support compiling Verilator with gcc/clang precompiled headers (#4579 )	2023-10-17 22:49:28 -04:00
Mariusz Glebocki	28bd7e5b19	Rework multithreading handling to separate by code units that use/never use it. (#4228 )	2023-09-24 22:12:23 -04:00
Adrien Le Masle	9cc218db3e	Fix incorrect multi-driven lint warning (#4231 ) (#4248 )	2023-06-01 08:43:17 -04:00
Wilson Snyder	5efe9367d2	Fix SystemC signal copy macro use (#4135 ).	2023-05-27 07:00:26 -04:00
Wilson Snyder	b24d7c83d3	Copyright year update	2023-01-01 10:18:39 -05:00
Jevin Sweval	299261714b	Fix crash in DFT due to width use after free (#3817 ) (#3820 )	2022-12-20 19:36:04 -05:00
Geza Lore	eaf09ba0e7	Dfg: resolve multi-driven signal ranges In order to avoid unexpected breakage on multi-driven variables, we resolve in DFG construction by using only the first driver encountered. Also issues the MULTIDRIVEN error for these signals.	2022-11-12 20:34:51 +00:00
Geza Lore	dbcaad99c5	Dfg: Fix crash on additional driver from non-DFG logic Ensure variables written by non-DFG code are kept Fixes #3740	2022-11-12 11:55:49 +00:00
Geza Lore	65e08f4dbf	Make all expressions derive from AstNodeExpr (#3721 ). Apart from the representational changes below, this patch renames AstNodeMath to AstNodeExpr, and AstCMath to AstCExpr. Now every expression (i.e.: those AstNodes that represent a [possibly void] value, with value being interpreted in a very general sense) has AstNodeExpr as a super class. This necessitates the introduction of an AstStmtExpr, which represents an expression in statement position, e.g : 'foo();' would be represented as AstStmtExpr(AstCCall(foo)). In exchange we can get rid of isStatement() in AstNodeStmt, which now really always represent a statement Peak memory consumption and verilation speed are not measurably changed. Partial step towards #3420	2022-11-03 16:02:16 +00:00
HungMingWu	196f3292d5	Improve V3Ast function usage ergonomics (#3650 ) Signed-off-by: HungMingWu <u9089000@gmail.com>	2022-10-21 14:12:12 +01:00
Geza Lore	90447d54d1	Make DfgConst hold V3Number directly Remove intermediary AstConst. No functional change intended.	2022-10-08 12:46:02 +01:00
Geza Lore	29a080dd9b	DFG: Special case representation of AstSel AstSel is a ternary node, but the 'widthp' is always constant and is hence redundant, and 'lsbp' is very often constant. As AstSel is fairly common, we special case as a DfgSel for the constant 'lsbp', and as 'DfgMux` for the non-constant 'lsbp'.	2022-10-06 19:59:01 +01:00
Geza Lore	965d99f1bc	DFG: Make implementation more similar to AST Use the same style, and reuse the bulk of astgen to generate DfgVertex related code. In particular allow for easier definition of custom DfgVertex sub-types that do not directly correspond to an AstNode sub-type. Also introduces specific names for the fixed arity vertices. No functional change intended.	2022-10-04 15:49:30 +01:00
Geza Lore	2a12b052f2	DFG: handle simple always blocks	2022-10-01 16:46:58 +01:00
Geza Lore	cc51966ad1	DFG: Remove unconneced variables early	2022-09-30 11:53:03 +01:00
Geza Lore	acebafcbc2	DFG: Partial support for unpacked arrays Representation and Ast / Dfg conversions available, for element-wise access only. Not much optimization yet (only CSE).	2022-09-29 19:00:45 +01:00
Geza Lore	1b17acdb01	DFG: Support AstSel and AstConcat on LHS of assignments Added DfgVertexVariadic to represent DFG vetices with a varying number of source operands. Converted DfgVar to be a variadic vertex, with each driver corresponding to a fixed range of bits in the packed variable. This allows us to handle AstSel on the LHS of assignments. Also added support for AstConcat on the LHS by selecting into the RHS as appropriate. This improves OpenTitan ST speed by ~13%	2022-09-26 19:54:52 +01:00
Geza Lore	9da012568c	Ensure DFG stats are consistent	2022-09-26 14:38:26 +01:00
Geza Lore	47bce4157d	Introduce DFG based combinational logic optimizer (#3527 ) Added a new data-flow graph (DFG) based combinational logic optimizer. The capabilities of this covers a combination of V3Const and V3Gate, but is also more capable of transforming combinational logic into simplified forms and more. This entail adding a new internal representation, `DfgGraph`, and appropriate `astToDfg` and `dfgToAst` conversion functions. The graph represents some of the combinational equations (~continuous assignments) in a module, and for the duration of the DFG passes, it takes over the role of AstModule. A bulk of the Dfg vertices represent expressions. These vertex classes, and the corresponding conversions to/from AST are mostly auto-generated by astgen, together with a DfgVVisitor that can be used for dynamic dispatch based on vertex (operation) types. The resulting combinational logic graph (a `DfgGraph`) is then optimized in various ways. Currently we perform common sub-expression elimination, variable inlining, and some specific peephole optimizations, but there is scope for more optimizations in the future using the same representation. The optimizer is run directly before and after inlining. The pre inline pass can operate on smaller graphs and hence converges faster, but still has a chance of substantially reducing the size of the logic on some designs, making inlining both faster and less memory intensive. The post inline pass can then optimize across the inlined module boundaries. No optimization is performed across a module boundary. For debugging purposes, each peephole optimization can be disabled individually via the -fno-dfg-peepnole-<OPT> option, where <OPT> is one of the optimizations listed in V3DfgPeephole.h, for example -fno-dfg-peephole-remove-not-not. The peephole patterns currently implemented were mostly picked based on the design that inspired this work, and on that design the optimizations yields ~30% single threaded speedup, and ~50% speedup on 4 threads. As you can imagine not having to haul around redundant combinational networks in the rest of the compilation pipeline also helps with memory consumption, and up to 30% peak memory usage of Verilator was observed on the same design. Gains on other arbitrary designs are smaller (and can be improved by analyzing those designs). For example OpenTitan gains between 1-15% speedup depending on build type.	2022-09-23 16:46:22 +01:00

45 Commits