nimbus-eth1

mirror of https://github.com/status-im/nimbus-eth1.git synced 2025-02-13 20:46:26 +00:00

Author	SHA1	Message	Date
Jordan Hrycaj	3c6400673d	Coredb fix kvt save only fringe condition (#2592 ) * Cosmetics, spelling, etc. * Aristo: make sure that a save cycle always commits even when empty why: If `Kvt` is tied to the `Aristo` DB save cycle, then this save cycle must also be committed if there is no data to save for `Aristo`. Otherwise this will lead to excessive core memory use with some fringe condition where Eth headers (or blocks) are downloaded while syncing and not really stored on disk. * CoreDb: Correct persistent save mode why: Saving `Kvt` first is seen as a harbinger (or canary) for `Aristo` as both run in sync. If `Kvt` succeeds saving first, so must be `Aristo` next. Other than this is a defect.	2024-09-04 13:48:38 +00:00
andri lim	4d9e288340	Wiring ForkedChainRef to other components (#2423 ) * Wiring ForkedChainRef to other components - Disable majority of hive simulators - Only enable pyspec_sim for the moment - The pyspec_sim is using a smaller RPC service wired to ForkedChainRef - The RPC service will gradually grow * Addressing PR review * Fix test_beacon/setup_env * Enable consensus_sim (#2441) * Enable consensus_sim * Remove isFile check * Enable Engine API jwt auth tests and exchange cap tests * Enable engine api in build_sim.sh * Wire ForkedChainRef to Engine API newPayload * Wire Engine API getBodies to ForkedChainRef * Wire Engine API api_forkchoice to ForkedChainRef * Wire more RPC methods to ForkedChainRef * Implement eth_syncing * Implement eth_call and eth_getlogs * TxPool: simplify smartHead * Fix smartHead usage * Fix txpool headDiff * Remove hasBlockHeader and use headerExists * Addressing review	2024-09-04 09:54:54 +00:00
Jacek Sieka	35cc78c86d	add metrics for rdb lru cache (#2586 ) This is a first step towards measuring the efficiency of the LRU caches over time - metrics can be collected during import or when running regulary. Since `nim-metrics` carries some overhead for its default way of reporting metrics, this PR implements a custom collector over atomic counters, given that this is one of the hottest spots in the block processing pipeline. Using a compile-time flag, the same metrics can be printed on exit which is useful when comparing different strategies for caching - here's a recent run over blocks 16000001-1616384 - this is a good candidate to expose in a better way in the future, maybe: ``` state vtype miss hit total hitrate Account Leaf 4909417 4466215 9375632 47.64% Account Branch 20742574 72015123 92757697 77.64% World Leaf 940483 1140946 2081429 54.82% World Branch 8224151 131496580 139720731 94.11% all all 34816625 209118864 243935489 85.73% ```	2024-09-02 17:34:10 +02:00
Jacek Sieka	ef1bab0802	avoid some trivial memory allocations (#2587 ) * pre-allocate `blobify` data and remove redundant error handling (cannot fail on correct data) * use threadvar for temporary storage when decoding rdb, avoiding closure env * speed up database walkers by avoiding many temporaries ~5% perf improvement on block import, 100x on database iteration (useful for building analysis tooling)	2024-09-02 16:03:10 +02:00
Jordan Hrycaj	a25ea63dec	Revert lazy implementation (#2585 )	2024-09-02 10:34:42 +00:00
Jacek Sieka	84a72c8658	Use zstd compression in bottommost layer (#2582 ) Tested up to block ~14m, zstd uses ~12% less space which seems to result in a small:ish (2-4%) performance improvement on block import speed - this seems like a better baseline for more extensive testing in the future. Pre: 57383308 kb Post: 50831236 kb	2024-08-30 17:32:13 +02:00
Jordan Hrycaj	42a08cfba9	Coredb and sync maintenance update (#2583 ) * bump metrics * Remove cruft * Cosmetics, update some logging, noise control * Renamed `CoreDb` function `hasKey` => `hasKeyRc` and provided `hasKey` why: Currently, `hasKey` returns a `Result[]` rather than a `bool` which is what one would expect from a function prototype of this name. This was a bit of an annoyance and cost unnecessary attention.	2024-08-30 11:18:36 +00:00
Jacek Sieka	8857fccb44	create per-fork opcode dispatcher (#2579 ) In the current VM opcode dispatcher, a two-level case statement is generated that first matches the opcode and then uses another nested case statement to select the actual implementation based on which fork it is, causing the dispatcher to grow by `O(opcodes) * O(forks)`. The fork does not change between instructions causing significant inefficiency for this approach - not only because it repeats the fork lookup but also because of code size bloat and missed optimizations. A second source of inefficiency in dispatching is the tracer code which in the vast majority of cases is disabled but nevertheless sees multiple conditionals being evaluated for each instruction only to remain disabled throughout exeuction. This PR rewrites the opcode dispatcher macro to generate a separate dispatcher for each fork and tracer setting and goes on to pick the right one at the start of the computation. This has many advantages: * much smaller dispatcher * easier to compile * better inlining * fewer pointlessly repeated instruction * simplified macro (!) * slow "low-compiler-memory" dispatcher code can be removed Net block import improvement at about 4-6% depending on the contract - synthetic EVM benchmnarks would show an even better result most likely.	2024-08-28 10:20:36 +02:00
web3-developer	8bf581e72d	Cleanup unused exp_getProofsByBlockNumber endpoint (#2577 ) * Cleanup unused exp_getProofsByBlockNumber endpoint.	2024-08-23 22:39:33 +08:00
Jacek Sieka	dbabe7e0a7	import: reduce stack usage (#2575 ) Because EthBlock is quite large, the stack usage that results from the multiple copies (temporary and not) present in the import command is larger than it should be - this PR moves some of that data to a closure environment allocated once per EthBlock - a larger restructuring of the code is due but in the meantime, this simple change speeds up garbage collection a little bit.	2024-08-22 10:06:45 +02:00
Jacek Sieka	d72a73de8b	avoid digest when loading era block (#2572 ) Computing the digest is unnecessary but takes a little bit of time - remove computation and reduce mem usage slightly when loading era blocks	2024-08-20 15:23:14 +02:00
Jordan Hrycaj	4db9c5c2d5	Small updates and fixes for rlpx suite (#2571 ) * Remove redundant `eth/68` message and clean up docu details: There is only eth/68 available at the moment * Allow to turn on chronicles line number logging in `Makefile` * Accept (and forget) tx hashes announcements why: Does no harm to just ignore it at the moment * Bump nim-eth (rlp fix)	2024-08-19 14:00:10 +00:00
Jacek Sieka	226cdb7c68	avoid exceptions, tx copy (#2569 ) * avoid some exceptions using the new compile-time strformat * avoid copying the full transaction only to normalise 2 fields (expensive)	2024-08-19 09:42:07 +02:00
Jacek Sieka	5941fef211	Avoid unnecessary layer copy (#2567 ) When the stack has an empty layer on top, there's no need to copy the contents of `top` to it since it would be the same. ~13% processing saved (!) pre ``` INF 2024-08-17 19:11:31.748+02:00 Imported blocks blockNumber=18667648 blocks=12000 importedSlot=7860043 txs=1797812 mgas=181135.177 bps=8.763 tps=1375.062 mgps=132.125 avgBps=6.798 avgTps=1018.501 avgMGps=102.617 elapsed=29m25s154ms ``` post ``` INF 2024-08-17 18:22:52.513+02:00 Imported blocks blockNumber=18667648 blocks=12000 importedSlot=7860043 txs=1797812 mgas=181135.177 bps=9.648 tps=1513.961 mgps=145.472 avgBps=7.876 avgTps=1179.998 avgMGps=118.888 elapsed=25m23s572ms ```	2024-08-19 08:46:04 +02:00
Jacek Sieka	43d93bcdab	Don't write slot hashes on import (#2564 ) The reverse slot hash mechanism causes quite a bit of database traffic but is broadly not useful except for iterating the storage of an account, something that a validator never does (it's used by the tracers). This flag adds one more thing that is not stored in the database, to be explored more comprehensively when designing full, validator and archive modes with different pruning options in the future. `ldb` says this is 60gb of data (!): ``` ldb --db=. --ignore_unknown_options --column_family=KvtGen approxsize --hex --from=0x05 --to=0x05ffffffffffffffffffffffffffffffffffffffffffffff 66488353954 ```	2024-08-16 08:22:51 +02:00
Jordan Hrycaj	4dbc1653ea	Cleanup (#2565 ) * Move snap un-dumpers to aristo unit test folder why: The only place where it is used, now to test the database against legacy snap sync dump samples. While the details of the dumped data have mostly outlived their purpuse, its use as entropy data thrown against `Aristo` has still been useful to find/debug tricky DB problems. * Remove cruft * `nimbus-eth1-blobs` not used anymore as test data source	2024-08-15 12:31:07 +00:00
Jordan Hrycaj	cbe5131927	Simplify aristo tree deletion functionality (#2563 ) * Cleaning up, removing cruft and debugging statements * Make `aristo_delta` fluffy compatible why: A sub-module that uses `chronicles` must import all possible modules used by a parent module that imports the sub-module. * update TODO	2024-08-14 12:09:30 +00:00
Jordan Hrycaj	d148de5b1c	Remove chunked rlpx (#2562 ) * bump nim-eth * Update make environment	2024-08-14 10:56:49 +00:00
Jordan Hrycaj	ce713d95fc	Aristo lazily delete larger subtrees (#2560 ) * Extract sub-tree deletion functions into separate sub-modules * Move/rename `aristo_desc.accLruSize` => `aristo_constants.ACC_LRU_SIZE` * Lazily delete sub-trees why: This gives some control of the memory used to keep the deleted vertices in the cached layers. For larger sub-trees, keys and vertices might be on the persistent backend to a large extend. This would pull an amount of extra information from the backend into the cached layer. For lazy deleting it is enough to remember sub-trees by a small set of (at most 16) sub-roots to be processed when storing persistent data. Marking the tree root deleted immediately allows to let most of the code base work as before. * Comments and cosmetics * No need to import all for `Aristo` here * Kludge to make `chronicle` usage in sub-modules work with `fluffy` why: That `fluffy` would not run with any logging in `core_deb` is a problem I have known for a while. Up to now, logging was only used for debugging. With the current `Aristo` PR, there are cases where logging might be wanted but this works only if `chronicles` runs without the `json[dynamic]` sinks. So this should be re-visited. * More of a kludge	2024-08-14 08:54:44 +00:00
Jordan Hrycaj	7becf4e389	Remove vertex ID recycle function (#2558 ) why: It is not safe in general to recycle vertex IDs while the `RocksDb` cache has `VertexID` rather than `RootedVertexID` where the former type seems preferable. In some fringe cases one might remove a vertex with key `(root1,vid)` and insert another vertex with key `(root2,vid)` while re-using the vertex ID `vid`. Without knowledge of `root1` and `root2`, the LRU cache will return the same vertex for `(root2,vid)` also for `(root1,vid)`.	2024-08-12 20:56:15 +00:00
Jacek Sieka	19451cadff	rebalance rocksdb cache sizes (#2557 ) Based on some simple testing done with a few combinations of cache sizes, it seems that the block cache has grown in importance compared to the where we were before changing on-disk format and adding a lot of other point caches. With these settings, there's roughly a 15% performance increase when processing blocks in the 18M range over the status quo while memory usage decreases by more than 1gb! Only a few values were tested so there's certainly more to do here but this change sets up a better baseline for any future optimizations. In particular, since the initial defaults were chosen root vertex id:s were introduced as key prefixes meaning that storage for each account will be grouped together and thus it becomes more likely that a block loaded from disk will be hit multiple times - this seems to give the block cache an edge over the row cache, specially when traversing the storage trie.	2024-08-12 05:52:09 +00:00
andri lim	b8e128203f	Rewire blockValue from Txpool to EngineAPI (#2554 )	2024-08-09 06:05:18 +07:00
Jacek Sieka	3dc30195ad	log http/jwt information on startup (#2553 )	2024-08-08 10:03:30 +00:00
Jacek Sieka	094486d0ce	Hash bump	2024-08-08 07:46:35 +02:00
Jacek Sieka	3cefd7ed38	move db init to init (#2552 ) When using the common interface, the database always (potentially) needs init - take the opportunity to log some basic database info on startup.	2024-08-08 07:45:30 +02:00
andri lim	d5786758b5	TxPool: Merge tx_chain and tx_packer to reduce complexity (#2549 ) * TxPool: Merge tx_chain and tx_packer to reduce complexity * Fix copyright year	2024-08-07 22:35:17 +07:00
Jordan Hrycaj	38572bd8ea	Cache a storage root ID forever in the leaf payload of an account (#2551 ) details: Stale root IDs are marked disabled while the ID is kept in the leaf payload. why: This might lead to further caching advantages.	2024-08-07 13:28:01 +00:00
Jordan Hrycaj	488bdbc267	Provide portal proof functionality with coredb (#2550 ) * Provide portal proof functions in `aristo_api` why: So it can be fully supported by `CoreDb` * Fix prototype in `kvt_api` * Fix node constructor for account leafs with storage trees * Provide simple path check based on portal proof functionality * Provide portal proof functionality in `CoreDb` * Update TODO list	2024-08-07 11:30:55 +00:00
andri lim	3cef119b78	Return empty list instead of error in getPooledTxs handler (#2547 )	2024-08-06 22:06:48 +07:00
Jordan Hrycaj	6bae929439	Added comments (#2546 )	2024-08-06 12:43:39 +00:00
Jordan Hrycaj	5b502a06c4	Added portal proof nodes generation functionality (#2539 ) * Extracted `test_tx.testTxMergeProofAndKvpList()` => separate file * Fix serialiser why: Typo lead to duplicate rlp-encoded nodes in chain * Remove cruft * Implemnt portal proof nodes generators `partXxxTwig()` * Add unit test for portal proof nodes generator `partAccountTwig()` * Cosmetics * Simplify serialiser return code format * Fix proof generator for extension nodes why: Code was simply bonkers, not detected before the unit tests were adapted to check for just this. * Implemented portal proof nodes verifier `partUntwig()` * Cosmetics * Fix `testutp` cli poblem	2024-08-06 11:29:26 +00:00
andri lim	ec118a438a	Refactor txpool: reduce complexity (#2542 )	2024-08-06 16:12:56 +07:00
andri lim	9dacfed943	Disable txpool in eth wire protocol handler (#2540 )	2024-08-06 11:26:55 +07:00
Jordan Hrycaj	01b5c08763	Revive json tracer unit tests (#2538 ) * Some `Aristo` clean-ups/updates * Re-implemented core-db tracer functionality * Rename nimbus tracer `no-tracer.nim` => `tracer.nim` why: Restore original name for easy diff tracking with upcoming update * Update nimbus tracer using new core-db tracer functionality * Updating json tracer unit tests * Enable json tracer unit tests	2024-08-01 10:41:20 +00:00
andri lim	e331c9e9b7	TxPool: Replace GasPrice and GasPriceEx with GasInt (#2537 ) * TxPool: Replace GasPrice and GasPriceEx with GasInt	2024-07-31 14:33:30 +07:00
Jordan Hrycaj	72c3ab8ced	Provide partial tree support for preloading tests (#2536 ) * Implement partial trees why: This is currently needed for unit tests to pre-load the database with test data similar to `proof` node pre-load. The basic features for `snap-sync` boundary proofs are available as well for future use. What is missing is the final proof verification and a complete storage data load/merge function (stub is available.) * Cosmetics, clean up	2024-07-29 20:15:17 +00:00
Jacek Sieka	bdc86b3fd4	small cleanups (#2526 ) * remove some redundant EH * avoid pessimising move (introduces a copy in this case!) * shift less data around when reading era files (reduces stack usage)	2024-07-26 12:32:01 +07:00
andri lim	254bda365f	Remove txpool sender locality (#2525 ) * Remove txpool sender locality We no longer distinct local or remote sender * Fix copyright year	2024-07-25 22:36:08 +07:00
andri lim	0cc730dd05	Fix CodeBytes: invalidPositions out of bound crash (#2523 )	2024-07-25 19:23:53 +07:00
andri lim	01ba18da74	Fix sepolia chain config: mergeForkBlock -> 1450409 (#2518 ) * Fix sepolia chain config: mergeForkBlock -> 1450407 * Fix test_forkid	2024-07-24 03:07:55 +00:00
Advaita Saha	08bbb0079f	faster slot finding in nimbus import (#2491 ) * faster slot finding in nimbus import * feat: blocknumber based slot finding * fix: formatting * added comments * fix: added is_execution_block * added comment	2024-07-22 21:17:07 +00:00
Jordan Hrycaj	1452e7b1c0	Misc updates (#2513 ) * Update config for Ledger and CoreDb why: Prepare for tracer which depends on the API jump table (as well as the profiler.) The API jump table is now enabled in unit/integration test mode piggybacking on the `unittest2DisableParamFiltering` compiler flag or on an extra compiler flag `dbjapi_enabled`. * No deed for error field in `NodeRef` why: Was opnly needed by proof nodes pre-loader which will be re-implemented * Cosmetics	2024-07-22 18:10:04 +00:00
andri lim	6d03acec30	TxPool refactoring: Simplify TxChainRef and remove gauges (#2506 ) This is one of the txPool refactoring series to make it ready for integration with the new ForkedChainRef	2024-07-19 16:24:36 +07:00
andri lim	fb196849ee	EVM cosmetic changes, one less indirect access of VmCpt (#2503 )	2024-07-19 08:44:01 +07:00
Jordan Hrycaj	5ac362fe6f	Aristo and kvt balancer management update (#2504 ) * Aristo: Merge `delta_siblings` module into `deltaPersistent()` * Aristo: Add `isEmpty()` for canonical checking whether a layer is empty * Aristo: Merge `LayerDeltaRef` into `LayerObj` why: No need to maintain nested object refs anymore. Previously the `LayerDeltaRef` object had a companion `LayerFinalRef` which held non-delta layer information. * Kvt: Merge `LayerDeltaRef` into `LayerRef` why: No need to maintain nested object refs (as with `Aristo`) * Kvt: Re-write balancer logic similar to `Aristo` why: Although `Kvt` was a cheap copy of `Aristo` it sort of got out of sync and the balancer code was wrong. * Update iterator over forked peers why: Yield additional field `isLast` indicating that the last iteration cycle was approached. * Optimise balancer calculation. why: One can often avoid providing a new object containing the merge of two layers for the balancer. This avoids copying tables. In some cases this is replaced by `hasKey()` look ups though. One uses one of the two to combine and merges the other into the first. Of course, this needs some checks for making sure that none of the components to merge is eventually shared with something else. * Fix copyright year	2024-07-18 21:32:32 +00:00
andri lim	ee323d5ff8	Optimize EVM stack usage (#2502 ) * EVM: Optimize CALL family stack usage * EVM: Optimize CREATE family stack usage * EVM: Optimize arith stack usage * EVM: Optimize stack usage in the rest of opcodes * Fix test_op_env and clean up unused imports * EVM: Optimize arithmetic binary ops	2024-07-18 18:59:53 +07:00
Jacek Sieka	df4a21c910	Store cached hash at the layer corresponding to the source data (#2492 ) When lazily verifying state roots, we may end up with an entire state without roots that gets computed for the whole database - in the current design, that would result in hashes for the entire trie being held in memory. Since the hash depends only on the data in the vertex, we can store it directly at the top-most level derived from the verticies it depends on - be that memory or database - this makes the memory usage broadly linear with respect to the already-existing in-memory change set stored in the layers. It also ensures that if we have multiple forks in memory, hashes get cached in the correct layer maximising reuse between forks. The same layer numbering scheme as elsewhere is reused, where -2 is the backend, -1 is the balancer, then 0+ is the top of the stack and stack. A downside of this approach is that we create many small batches - a future improvement could be to collect all such writes in a single batch, though the memory profile of this approach should be examined first (where is the batch kept, exactly?).	2024-07-18 09:13:56 +02:00
Jordan Hrycaj	6677f57ea9	Aristo balancer clean up (#2501 ) * Remove `chunkedMpt` from `persistent()`/`stow()` function why: Proof-mode code was removed with PR #2445 and needs to be re-designed. * Remove unused `beStateRoot` argument from `deltaMerge()` * Update/drastically simplify `txStow()` why: Got rid of many boundary conditions details: Many pre-conditions have changed. In particular, previous versions used the account state (hash) which was conveniently available and checked it against the backend in order to find out whether there was something to do, at all. Currently, only an empty set of all tables in the delta layer has the balancer update ignored. Notable changes are: * no check against account state (see above) * balancer filters have no hash signature (some legacy stuff left over from journals) * no (shap sync) proof data which made the generation of the a top layer more complex * Cosmetics, cruft removal * Update unit test file & function name why: Was legacy module	2024-07-17 19:27:33 +00:00
andri lim	cfe14f1825	EVM: use assign2 whenever possible (#2499 ) Before: GST finish in 59 secs. After: GST finish in 52 secs!	2024-07-17 20:48:50 +07:00
andri lim	8d1e21bbae	Simplify txPool gasLimit calculator (#2498 ) Our need is only a baseline tx pool gasLimit calculator. If need we can expand it in the future. But for now, a simple but understandable tx pool is more important.	2024-07-17 20:48:35 +07:00

1 2 3 4 5 ...

1884 Commits