forks/binaryen.git -

	Commit message (Collapse)	Author	Age	Files	Lines
*	[Strings] Fix non-nullable string emitting in the binary format (#5756)	Alon Zakai	2023-06-07	1	-3/+9
\| \| \|	Related to #5737 which did something similar for other types.
*	[NFC] Remove redundant code from EffectAnalyzer (#5754)	Alon Zakai	2023-06-06	1	-5/+0
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This PR removes a check for transfersControlFlow() && other.trap which is already checked higher up in the code, here: transfersControlFlow() && other.hasSideEffects() In binaryen/src/ir/effects.h , lines 223 to 224 That last code handles the first code because trapping is part of hasSideEffects().
*	Move casts which are immediate children of local.gets to earlier local.gets ↵	Bruce He	2023-06-06	1	-16/+324
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	(#5744) In the OptimizeCasts pass, it is useful to move more refined casts as early as possible without causing side-effects. This will allow such casts to potentially trap earlier, and will allow the OptimizeCasts pass to use more refined casts earlier. This change allows a more refined cast to be duplicated at an earlier local.get expression. The later instance of the cast will then be eliminated in a later optimization pass. For example, if we have the following instructions: (drop (local.get $x) ) (drop (ref.cast $A (local.get $x) ) (drop (ref.cast $B (local.get $x) ) ) Where $B is a sublcass of $A, we can convert this to: (drop (ref.cast $B (local.get $x) ) ) (drop (ref.cast $A (local.get $x) ) (drop (ref.cast $B (local.get $x) ) ) Concretely we will save the first cast to a local and use it in the other local.gets.
*	Fix emitting of function reference types without GC (#5737)	Thomas Lively	2023-06-05	1	-14/+17
\| \| \| \| \| \| \| \| \|	We previously had logic to emit GC types used in the IR as their corresponding top types when GC was not enabled (so e.g. nullfuncref would be emitted as funcref), but the logic was not robust enough and non-null function references were not properly emitted as funcref. Refactor the relevant code to be more robust and future-proof, and add a test demonstrating that the lowering works as intended.
*	StackIR: Remove nops (#5746)	Alon Zakai	2023-05-30	1	-0/+17
\| \| \| \| \| \| \|	No nop instruction is necessary in wasm, so in StackIR we can simply remove them all. Fixes #5745
*	wasm-merge: Preserve imports when copying module items (#5743)	Jérôme Vouillon	2023-05-26	1	-0/+4
\| \| \| \|	The import information of Tags and Memories was not preserved.
*	Revert "Update br_on_cast binary and text format (#5734)" (#5740)	Alon Zakai	2023-05-23	6	-40/+46
\| \| \| \| \| \| \|	This reverts commit b7b1d0df29df14634d2c680d1d2c351b624b4fbb. See comment at the end of #5734: It turns out that dropping the old opcodes causes problems for current users, so let's revert this for now, and later we can figure out how best to do the update.
*	Fuzzer: Limit ArrayNew sizes most of the time (#5738)	Alon Zakai	2023-05-22	1	-2/+11
\|
*	TypeSSA: Handle collisions by adding a hash to ensure a fresh rec group (#5724)	Alon Zakai	2023-05-19	2	-22/+100
\| \| \|	Fixes #5720
*	Update br_on_cast binary and text format (#5734)	Thomas Lively	2023-05-19	6	-46/+40
\| \| \| \| \| \| \| \| \| \|	The final versions of the br_on_cast and br_on_cast_fail instructions have two reference type annotations: one for the input type and one for the cast target type. In the binary format, this is represented as a flags byte followed by two encoded heap types. Since these instructions have been in flux for a while, do not attempt to maintain backward compatibility with older versions of the instructions. Instead, upgrade all of the tests at once to use the new versions of the instructions. Drop some binary tests of deprecated instruction encodings that would be more effort to update than they're worth.
*	Improve TODO docs in OptimizeCasts (#5732)	Alon Zakai	2023-05-18	1	-7/+17
\|
*	Vacuum code leading up to a trap in TrapsNeverHappen mode (#5228)	Alon Zakai	2023-05-17	1	-1/+79
\| \| \| \| \| \| \| \| \| \| \| \|	This adds two rules to vacuum in TNH mode: if (..) trap() => if (..) {} { stuff, trap() } => {} That is, we assume traps never happen so an if will not branch to one, and code right before a trap can be assumed to not execute. Together, we should be removing practically all possible code in TNH mode (though we could also add support for br_if etc.).
*	avoid incomplete type in a vector (#5730)	walkingeyerobot	2023-05-17	1	-17/+17
\| \| \| \| \|	Swap the order of struct declarations to avoid using an incomplete type in a vector. In C++20, using std::vector with an incomplete type often becomes a build failure due to increased usage of constexpr for vector members.
*	Print function types on function imports in the text format (#5727)	Alon Zakai	2023-05-17	1	-0/+4
\| \| \| \|	The function type should be printed there just like for non-imported functions.
*	EffectAnalyzer: Do not clear break targets before walk()/visit() (#5723)	Alon Zakai	2023-05-17	1	-7/+0
\| \| \| \| \| \|	We depend on repeated calls to walk/visit accumulating effects, so this was a bug; if we want to clear stuff then we create a new EffectAnalyzer. Removing that fixes the attached testcase. Also added a unit test.
*	Reintroduce wasm-merge (#5709)	Alon Zakai	2023-05-16	3	-6/+631
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	We used to have a wasm-merge tool but removed it for a lack of use cases. Recently use cases have been showing up in the wasm GC space and elsewhere, as people are using more diverse toolchains together, for example a project might build some C++ code alongside some wasm GC code. Merging those wasm files together can allow for nice optimizations like inlining and better DCE etc., so it makes sense to have a tool for merging. Background: * Removal: #1969 * Requests: * wasm-merge - why it has been deleted #2174 * Compiling and linking wat files #2276 * wasm-link? #2767 This PR is a compete rewrite of wasm-merge, not a restoration of the original codebase. The original code was quite messy (my fault), and also, since then we've added multi-memory and multi-table which makes things a lot simpler. The linking semantics are as described in the "wasm-link" issue #2767 : all we do is merge normal wasm files together and connect imports and export. That is, we have a graph of modules and their names, and each import to a module name can be resolved to that module. Basically, like a JS bundler would do for JS, or, in other words, we do the same operations as JS code would do to glue wasm modules together at runtime, but at compile time. See the README update in this PR for a concrete example. There are no plans to do more than that simple bundling, so this should not really overlap with wasm-ld's use cases. This should be fairly fast as it works in linear time on the total input code. However, it won't be as fast as wasm-ld, of course, as it does build Binaryen IR for each module. An advantage to working on Binaryen IR is that we can easily do some global DCE after merging, and further optimizations are possible later.
*	[NFC] Optimize ArrayNew zero construction (#5722)	Alon Zakai	2023-05-15	1	-1/+2
\| \| \| \| \| \| \| \| \| \|	All array elements have the same type, so we can construct a single zero and just copy it. This makes ArrayNew of large arrays 2x faster. I also experimented with putting Literal::makeZero in a header, in hopes of inlining leading to licm helping here, but that did not help at all unfortunately, at least not in gcc.
*	[Strings] Adopt new instruction binary encoding (#5714)	Jérôme Vouillon	2023-05-12	10	-118/+129
\| \| \| \| \| \| \| \| \| \| \|	See WebAssembly/stringref#46. This format is already adopted by V8: https://chromium-review.googlesource.com/c/v8/v8/+/3892695. The text format is left unchanged (see #5607 for a discussion on the subject). I have also added support for string.encode_lossy_utf8 and string.encode_lossy_utf8 array (by allowing the replace policy for Binaryen's string.encode_wtf8 instruction).
*	[analysis] Add a new iterable CFG utility (#5712)	Thomas Lively	2023-05-12	6	-0/+351
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Add a new "analysis" source directory that will contain the source for a new static program analysis framework. To start the framework, add a CFG utility that provides convenient iterators for iterating through the basic blocks of the CFG as well as the predecessors, successors, and contents of each block. The new CFGs are constructed using the existing CFGWalker, but they are different in that the new utility is meant to provide a usable representation of a CFG whereas CFGWalker is meant to allow collecting arbitrary information about each basic block in a CFG. For testing and debugging purposes, add `print` methods to CFGs and basic blocks. This requires exposing the ability to print expression contents excluding children, which was something we previously did only for StackIR. Also add a new gtest file with a test for constructing and printing a CFG. The test reveals some strange properties of the current CFG construction, including empty blocks and strange placement of `loop` instructions, but fixing these problems is left as future work.
*	Extend drop.h and use it in Directize (#5713)	Alon Zakai	2023-05-10	3	-53/+48
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This adds an option to ignore effects in the parent in getDroppedChildrenAndAppend. With that, this becomes usable in more places, like Directize, basically in situations where we know we can ignore effects in the parent (since we've inferred they are not needed). This lets us get rid of some boilerplate code in Directize. Diff without whitespace is a lot smaller. A large other part of the diff is a rename of curr => parent which I think it makes it more readable as then parent/children is a clear contrast, and then the new parameter "ignore/ notice parent effects" is obviously connected to "parent". The top comment in drop.cpp is removed as it just duplicated the top comment in the header drop.h. This is basically NFC but using drop.h does bring the advantage of emitting less code, see the test changes, so it is noticeable in the IR. This is a refactoring PR in preparation for a larger improvement to Directize that will also benefit from this new drop capability.
*	Gate all partial inlining behind the partial-inlining-ifs flag (#5710)	Alon Zakai	2023-05-10	1	-6/+10
\| \| \| \| \| \| \| \| \|	#4191 meant to do that, I think, but only did so for "pattern B". This does it for all patterns, and adds assertions. In theory this could regress code that benefits from partial inlining of "pattern A" (since this PR stops doing it by default), but I did not see a significant difference on any benchmarks, and it is easy to re-enable that behavior by doing --partial-inlining-ifs=1.
*	Add a "mayNotReturn" effect (#5711)	Alon Zakai	2023-05-10	1	-11/+5
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This changes loops from having the effect "may trap (timeout)" to having "may not return." The only noticeable difference is in TrapsNeverHappen mode, which ignores the former but not the latter. So after this PR, in TNH mode we do not optimize away an infinite loop that seems to have no other side effects. We may also use this for other things in the future, like continuations/stack switching. There are upsides and downsides to allowing the optimizer to remove infinite loops (the C and C++ communities have had interesting discussions on that topic over the years...) but it seems safer to not optimize them out for now, to let the most code work properly. If a need comes up to optimize such code, we can look at various options then (like a flag to ignore infinite loops). See discussion in #5228
*	Remove TypeUpdater in Vacuum and always ReFinalize (#5707)	Alon Zakai	2023-05-09	1	-48/+1
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	TypeUpdater::remove must be called after removing a thing from the tree. If not, then we can get confused by something like this: (block $b (br $b) ) If we first call TypeUpdater::remove then we see that the block's only br is going away, so it becomes unreachable. But when we then remove the br then the block should have type none. Removing the br first from the IR, and then calling TypeUpdater::remove, is the safe way to do it. However, changing that order in Vacuum is not trivial. After looking into this, I realized that it is just simpler to remove TypeUpdater entirely. Instead, we can ReFinalize at the end unconditionally. This has the downside that we do not propagate type updates "as we go", but that should be very rare. Another downside is that TypeUpdater tracks the number of brs, which can help remove code like in the one test that regresses here (see comment there). But I'm not sure that removal was valid - Vacuum should not really be doing it, and it looks like its related to this bug actually. Instead, we have a dedicated pass for removing unused brs - RemoveUnusedBrs - so leave things for it. This PR's benefit, aside from now handling the fuzz testcase, is that it makes the code simpler and faster. I see a 10-25% speedup on the Vacuum pass on various binaries I tested on. (Vacuum is one of our faster passes anyhow, though, so the runtime of -O1 is not much improved.) Another minor benefit might be that running ReFinalize more often can propagate type info more quickly, thanks to #5704 etc. But that is probably very minor.
*	Fix optimizeAddedConstants on GC-introduced unreachability (#5706)	Alon Zakai	2023-05-09	1	-3/+8
\|
*	[Wasm GC] wasm-ctor-eval: Handle cycles of data (#5685)	Alon Zakai	2023-05-05	2	-57/+377
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	A cycle of data is something we can't just naively emit as wasm globals. If at runtime we end up, for example, with an object A that refers to itself, then we can't just emit (global $A (struct.new $A (global.get $A))) The struct.get is of this very global, and such a self-reference is invalid. So we need to break such cycles as we emit them. The simple idea used here is to find paths in the cycle that are nullable and mutable, and replace the initial value with a null that is fixed up later in the start function: (global $A (struct.new $A (ref.null $A))) (func $start (struct.set (global.get $A) (global.get $A))) ) This is not optimal in terms of breaking cycles, but it is fast (linear time) and simple, and does well in practice on j2wasm (where cycles in fact occur).
*	[NFC] Add some comments for split inlining crossing the normal limit (#5702)	Alon Zakai	2023-05-05	1	-3/+46
\|
*	Generate unique block names when inlining (#5697)	Alon Zakai	2023-05-05	1	-3/+16
\| \| \| \| \| \| \| \| \| \| \|	Each time we inline we put the contents in a block. Before we used the same name each time we inlined the same method, and as a result had many conflicts if a function was inlined many times. With this PR we emit a different name each time. This is not 100% NFC as it does change block names, which is observable in the IR (as can be seen in the test updates). This helps #5696 in speeding up UniqueNameManner.
*	[Wasm GC] Automatically make RefCast heap types more precise (#5704)	Alon Zakai	2023-05-05	1	-1/+15
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	We already did this for nullablilty, and so for the same reasons we should do it for heap types as well. Also, I realized that doing so would solve #5703, which is the new test added for TypeRefining here. The fuzz bug solved here is that our analysis of struct gets/sets will skip copy operations - a read from a field that is written into it. And we skip fallthrough values while doing so, since it doesn't matter if the read goes through an if arm or a cast. An if would automatically get a more precise type during refinalize, so this PR does the same for a cast basically. Fixes #5703
*	[NFC] Scan first in UniqueNameMapper (#5696)	Alon Zakai	2023-05-05	1	-2/+43
\| \| \| \| \| \| \| \|	We can scan far faster than the name mapping process itself, so do that first, as in the common case nothing needs to be fixed up. This together with #5697 significantly reduces the overhead in --inlining-optimizing of UniqueNameMapper, from something like 50% to 10%.
*	[NFC] Track the kinds of items that names refer to in ↵	Alon Zakai	2023-05-05	5	-48/+135
\| \| \| \| \| \| \| \| \| \| \| \|	wasm-delegations-fields (#5690) This makes delegations-fields track Kinds. That is, rather than say a field is just a Name, we can say it is a name of kind Function. This allows users to track references to functions, tables, memories, etc., in a simple and generic way, avoiding duplicated code which we have atm. (In particular this will help wasm-merge in the future.) This also uses that functionality in two small places to show the benefits (see memory-utils.cpp and MemoryPacking.cpp).
*	[NFC] Refactor each of ArrayNewSeg and ArrayInit into subclasses for ↵	Alon Zakai	2023-05-04	25	-368/+563
\| \| \| \| \| \| \| \| \| \| \|	Data/Elem (#5692) ArrayNewSeg => ArrayNewSegData, ArrayNewSegElem ArrayInit => ArrayInitData, ArrayInitElem Basically we remove the opcode and use the class type to differentiate them. This adds some code but it makes the representation simpler and more compact in memory, and it will help with #5690
*	Fix DeadArgumentElimination return value opts on nesting+recursion (#5701)	Alon Zakai	2023-05-04	1	-13/+16
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The testcase here has a recursive call that is also nested in itself, something like (return (call $me (return .. This found a bug in our return value removal logic. When we remove a return value we both modify call sites (to add drops) and modify returns (to remove their values). One of those uses pointers into the IR which the other invalidated, so the order of the two matters. This PR just reorders that code to fix the bug.
*	Fallback to direct inlining if the outline will be inlined. (#5698)	Goktug Gokdogan	2023-05-04	1	-4/+38
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This workarounds the extra work around the edge case where; - Function is too big to full-inline - It is a candidate for partial inline - Outlined version becomes eligible for full-inline. In such a case, binaryen would introduce a temporary state with partial inlined functions and later on inline them. J2CL hit this scenario for String literal which resulted in significant regressions in compilation time. This patch updates partial inlining analysis to identify the edge case and direct to full-inlining when that happens.
*	Emit memory segment index for data segments (#5699)	Alon Zakai	2023-05-03	1	-0/+9
\| \| \|	Before this fix we would flip all data segments to use the first memory.
*	[NFC] Start tracking InliningMode and memoize it to avoid re-analysis. (#5695)	Goktug Gokdogan	2023-05-02	1	-73/+72
\|
*	[Wasm GC] Always refinalize in SignatureRefining (#5694)	Alon Zakai	2023-05-01	1	-9/+2
\| \| \| \|	We used to refine only for result changes, but param changes can also lead to opportunities.
*	[NFC] Inlining: Split maybeSplit into canSplit/doSplit (#5691)	Goktug Gokdogan	2023-04-28	1	-115/+108
\| \| \|	This will make future improvements easier.
*	Fix name deduplication with partial names sections (#5689)	Alon Zakai	2023-04-28	1	-0/+43
\| \| \| \| \| \| \| \| \| \|	We already deduplicated names in the names section (to defend against a weird binary), but we also need to deduplicate the names of items not in the names section, so they don't overlap with the names that are. See example in the testcase. Normally wasm files use names for all items in each group. This only became noticeable in some wasm-ctor-eval work where new temp globals were added that were not given names.
*	[NFC] Assert that module maps are the right size (#5687)	Alon Zakai	2023-04-25	1	-0/+8
\| \| \| \|	If the names are not unique then the map would be smaller than the vector it is built from.
*	[Wasm GC] Ignore GC cycle leaks in LSan (#5686)	Alon Zakai	2023-04-24	1	-9/+26
\| \| \| \| \|	Leaks happen since we use std::shared_ptr which does not handle cycles. But since Binaryen isn't used in long-running code it's probably find to just let them leak, and ignore them in LSan, for now.
*	[EH] Support assert_exception (#5684)	Heejin Ahn	2023-04-23	1	-3/+19
\| \| \| \| \| \| \| \|	`assert_exception` is similar to `assert_trap` but for exceptions, which is supported in the interpreter of the EH proposal (https://github.com/WebAssembly/exception-handling/tree/main/interpreter). We've been using `assert_trap` for both traps and exceptions, but this PR distinguishes them.
*	[NFC] Simplify rec group initialization in wasm-type.cpp (#5683)	Thomas Lively	2023-04-20	2	-34/+28
\| \| \| \| \| \| \| \|	Now that we no longer support constructing basic heap types in TypeBuilder, we can fully initialize rec groups when they are created, rather than having to initialize them later during the build step after any basic types have been canonicalized. Alongside that change, also simplify the process of initializing a type builder slot to avoid completely overwriting the HeapTypeInfo in the slot and avoid the hacky workarounds that required.
*	[Wasm GC] ReFinalize when needed in SimplifyGlobals (#5682)	Alon Zakai	2023-04-20	1	-4/+23
\|
*	[Wasm GC] Fix a trapsNeverHappen corner case with if/select of a trapping ↵	Alon Zakai	2023-04-20	1	-4/+22
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	arm (#5681) The logic says that if an if/select has an arm that returns a null type, and the if/select goes into a cast, then we can ignore that arm in tnh mode (as it would trap, and we are ignoring the possibility of a trap). But it is not enough to return a null type - the null must actually flow out, rather than say a return be executed before. One existing test needed adjustment, as it used calls for "thing with effects". But a call can transfer control flow when EH is enabled, and this pass has -all. Rather than mess with the features, I switched the effects to be locals.
*	[NFC] Minor simplifications in wasm-type.cpp (#5680)	Thomas Lively	2023-04-20	1	-43/+12
\| \| \| \| \| \| \| \| \| \| \|	This capability was originally introduced to support calculating LUBs in the equirecursive type system, but has not been needed for anything except tests since the equirecursive type system was removed. Since building basic heap types is no longer useful and was a source of significant complexity, remove the APIs that allowed it and the tests that used those APIs. Also remove test/example/type-builder.cpp, since a significant portion of it tested the removed APIs and the rest is already better tested in test/gtest/type-builder.cpp.
*	Remove the ability to construct basic types in a TypeBuilder (#5678)	Thomas Lively	2023-04-19	6	-274/+35
\| \| \| \| \| \| \| \| \| \| \|	This capability was originally introduced to support calculating LUBs in the equirecursive type system, but has not been needed for anything except tests since the equirecursive type system was removed. Since building basic heap types is no longer useful and was a source of significant complexity, remove the APIs that allowed it and the tests that used those APIs. Also remove test/example/type-builder.cpp, since a significant portion of it tested the removed APIs and the rest is already better tested in test/gtest/type-builder.cpp.
*	Disable the memory64 feature in Memory64Lowering.cpp (#5679)	Thomas Lively	2023-04-19	1	-0/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	* Disable sign extension in SignExtLowering.cpp The sign extension lowering pass would previously lower away the sign extension instructions, but it wouldn't disable the sign extension feature, so follow-on passes such as optimize-instructions could reintroduce sign extension instructions. Fix the pass to disable the sign extension feature to prevent sign extension instructions from being reintroduced later. * update pass description * Disable the memory64 feature in Memory64Lowering.cpp For consistency with other feature lowering passes, disable memory64 in addition to lowering its use away. Although no other passes would introduce new uses of memory64 at the moment, this makes the lowering pass more robust against a future where memory64 might accidentally be reintroduced after being lowered away. * Update test/lit/passes/memory64-lowering-features.wast Co-authored-by: Alon Zakai <azakai@google.com> --------- Co-authored-by: Alon Zakai <azakai@google.com>
*	Disable sign extension in SignExtLowering.cpp (#5676)	Thomas Lively	2023-04-19	2	-1/+10
\| \| \| \| \| \| \| \| \| \| \| \| \|	* Disable sign extension in SignExtLowering.cpp The sign extension lowering pass would previously lower away the sign extension instructions, but it wouldn't disable the sign extension feature, so follow-on passes such as optimize-instructions could reintroduce sign extension instructions. Fix the pass to disable the sign extension feature to prevent sign extension instructions from being reintroduced later. * update pass description
*	Fuzzer: Use subtype consistently in make() (#5674)	Alon Zakai	2023-04-19	1	-4/+4
\|
*	[Wasm GC] Fix GUFA on array.init of a bottom type (#5675)	Alon Zakai	2023-04-19	1	-2/+6
\|