go-ethereum

Commit Graph

Author	SHA1	Message	Date
rjl493456442	eff0bed91b	core/rawdb: freezer index repair (#29792 ) This pull request removes the `fsync` of index files in freezer.ModifyAncients function for performance gain. Originally, fsync is added after each freezer write operation to ensure the written data is truly transferred into disk. Unfortunately, it turns out `fsync` can be relatively slow, especially on macOS (see https://github.com/ethereum/go-ethereum/issues/28754 for more information). In this pull request, fsync for index file is removed as it turns out index file can be recovered even after a unclean shutdown. But fsync for data file is still kept, as we have no meaningful way to validate the data correctness after unclean shutdown. --- But why do we need the `fsync` in the first place? As it's necessary for freezer to survive/recover after the machine crash (e.g. power failure). In linux, whenever the file write is performed, the file metadata update and data update are not necessarily performed at the same time. Typically, the metadata will be flushed/journalled ahead of the file data. Therefore, we make the pessimistic assumption that the file is first extended with invalid "garbage" data (normally zero bytes) and that afterwards the correct data replaces the garbage. We have observed that the index file of the freezer often contain garbage entry with zero value (filenumber = 0, offset = 0) after a machine power failure. It proves that the index file is extended without the data being flushed. And this corruption can destroy the whole freezer data eventually. Performing fsync after each write operation can reduce the time window for data to be transferred to the disk and ensure the correctness of the data in the disk to the greatest extent. --- How can we maintain this guarantee without relying on fsync? Because the items in the index file are strictly in order, we can leverage this characteristic to detect the corruption and truncate them when freezer is opened. Specifically these validation rules are performed for each index file: For two consecutive index items: - If their file numbers are the same, then the offset of the latter one MUST not be less than that of the former. - If the file number of the latter one is equal to that of the former plus one, then the offset of the latter one MUST not be 0. - If their file numbers are not equal, and the latter's file number is not equal to the former plus 1, the latter one is valid And also, for the first non-head item, it must refer to the earliest data file, or the next file if the earliest file is not sufficient to place the first item(very special case, only theoretical possible in tests) With these validation rules, we can detect the invalid item in index file with greatest possibility. --- But unfortunately, these scenarios are not covered and could still lead to a freezer corruption if it occurs: All items in index file are in zero value It's impossible to distinguish if they are truly zero (e.g. all the data entries maintained in freezer are zero size) or just the garbage left by OS. In this case, these index items will be kept by truncating the entire data file, namely the freezer is corrupted. However, we can consider that the probability of this situation occurring is quite low, and even if it occurs, the freezer can be considered to be close to an empty state. Rerun the state sync should be acceptable. Index file is integral while relative data file is corrupted It might be possible the data file is corrupted whose file size is extended correctly with garbage filled (e.g. zero bytes). In this case, it's impossible to detect the corruption by index validation. We can either choose to `fsync` the data file, or blindly believe that if index file is integral then the data file could be integral with very high chance. In this pull request, the first option is taken.	1 month ago
wangyifan	cd58897f18	core/rawdb: implement size reporting for live items in freezer_table (#28525 ) This is the fix to issue #27483. A new hiddenBytes() is introduced to calculate the byte size of hidden items in the freezer table. When reporting the size of the freezer table, size of the hidden items will be subtracted from the total size. --------- Co-authored-by: Yifan <Yifan Wang> Co-authored-by: Gary Rong <garyrong0905@gmail.com>	11 months ago
rjl493456442	326fa00759	core/rawdb: fsync the index file after each freezer write (#28483 ) * core/rawdb: fsync the index and data file after each freezer write * core/rawdb: fsync the data file in freezer after write	1 year ago
rjl493456442	3853f50082	trie/triedb/pathdb, core/rawdb: enhance error message in freezer (#28198 ) This PR adds more error message for debugging purpose.	1 year ago
Delweng	545f4c5547	core/rawdb: no need to run truncateFile for readonly mode (#28145 ) Avoid truncating files, if ancients are opened in readonly mode. With this change, we return error instead of trying (and failing) to repair	1 year ago
rjl493456442	0b1f97e151	core/rawdb: support freezer batch read with no size limit (#27687 ) This change adds the ability to perform reads from freezer without size limitation. This can be useful in cases where callers are certain that out-of-memory will not happen (e.g. reading only a few elements). The previous API was designed to behave both optimally and secure while servicing a request from a peer, whereas this change should _not_ be used when an untrusted peer can influence the query size.	1 year ago
Delweng	7f3fc15a8b	core/rawdb: update freezertable read meter (#26946 ) The meter for "for measuring the effective amount of data read" within the freezertable was never updated. This change remedies that. --------- Signed-off-by: jsvisa <delweng@gmail.com>	2 years ago
s7v7nislands	905a723fae	core/rawdb: use atomic int added in go1.19 (#26935 )	2 years ago
Martin Holst Swende	0b53b29078	core/rawdb: fix cornercase shutdown behaviour in freezer (#26485 ) This PR does a few things. It fixes a shutdown-order flaw in the chainfreezer. Previously, the chain-freezer would shutdown the freezer backend first, and then signal for the loop to exit. This can lead to a scenario where the freezer tries to fsync closed files, which is an error-conditon that could lead to exit via log.Crit. It also makes the printout more detailed when truncating 'dangling' items, by showing the exact number instead of approximate MB. This PR also adds calls to fsync files before closing them, and also makes the `db inspect` command slightly more robust.	2 years ago
rjl493456442	e04d63ebd3	core/rawdb: fsync head data file before closing it (#26490 ) This PR fixes an issue which might result in data lost in freezer. Whenever mutation happens in freezer, all data will be written into head data file and it will be rotated with a new one in case the size of file reaches the threshold. Theoretically, the rotated old data file should be fsync'd to prevent data loss. In freezer.Sync function, we only fsync: (1) index file (2) meta file and (3) head data file. So this PR forcibly fsync the head data file if mutation happens in the boundary of data file.	2 years ago
Felix Lange	193f350eb9	core/rawdb: improve freezerTable.Sync (#26245 ) While investigating #22374, I noticed that the Sync operation of the freezer does not take the table lock. It also doesn't call sync for all files if there is an error with one of them. I doubt this will fix anything, but didn't want to drop the fix on the floor either.	2 years ago
rjl493456442	b9ba6f6e4d	core/rawdb: open meta file in read only mode (#26009 )	2 years ago
Justin Traglia	2c5648d891	all: fix some typos (#25551 ) * Fix some typos * Fix some mistakes * Revert 4byte.json * Fix an incorrect fix * Change files to fails	2 years ago
rjl493456442	e44d6551c3	cmd, core, ethdb, node: move chain freezer one folder deeper (#25487 ) * cmd, core, ethdb, node: create chain freezer in a sub folder * core/rawdb: remove unused code * core, ethdb, node: add AncientDatadir API back * cmd, core: extend freezer info dump for sub-ancient-store * core/rawdb: rework freezer inspector * core/rawdb: address comments from Peter * core/rawdb: fix build issue	2 years ago
s7v7nislands	4024c1e869	fix typo (#24731 )	3 years ago
rjl493456442	6afb717be5	core/rawdb: fix db commands (#24540 )	3 years ago
rjl493456442	538a868384	core/rawdb, cmd, ethdb, eth: implement freezer tail deletion (#23954 ) * core/rawdb, cmd, ethdb, eth: implement freezer tail deletion * core/rawdb: address comments from martin and sina * core/rawdb: fixes cornercase in tail deletion * core/rawdb: separate metadata into a standalone file * core/rawdb: remove unused code * core/rawdb: add random test * core/rawdb: polish code * core/rawdb: fsync meta file before manipulating the index * core/rawdb: fix typo * core/rawdb: address comments	3 years ago
Sina Mahmoodi	4aab440ee2	core/rawdb: enforce readonly in freezer instantiation (#24119 ) * freezer: add readonly flag to table * freezer: enforce readonly in table repair * freezer: enforce readonly in newFreezer * minor fix * minor * core/rawdb: test that writing during readonly fails * rm unused log * check readonly on batch append * minor * Revert "check readonly on batch append" This reverts commit `2ddb5ec4ba`. * review fixes * minor test refactor * attempt at fixing windows issue * add comment re windows sync issue * k->kind * open readonly db for genesis check Co-authored-by: Martin Holst Swende <martin@swende.se>	3 years ago
Martin Holst Swende	794c6133ef	core/rawdb: freezer batch write (#23462 ) This change is a rewrite of the freezer code. When writing ancient chain data to the freezer, the previous version first encoded each individual item to a temporary buffer, then wrote the buffer. For small item sizes (for example, in the block hash freezer table), this strategy causes a lot of system calls for writing tiny chunks of data. It also allocated a lot of temporary []byte buffers. In the new version, we instead encode multiple items into a re-useable batch buffer, which is then written to the file all at once. This avoids performing a system call for every inserted item. To make the internal batching work, the ancient database API had to be changed. While integrating this new API in BlockChain.InsertReceiptChain, additional optimizations were also added there. Co-authored-by: Felix Lange <fjl@twurst.com>	3 years ago
Martin Holst Swende	5f98020a21	core/rawdb: implement sequential reads in freezer_table (#23117 ) * core/rawdb: implement sequential reads in freezer_table * core/rawdb, ethdb: add sequential reader to db interface * core/rawdb: lint nitpicks * core/rawdb: fix some nitpicks * core/rawdb: fix flaw with deferred reads not being performed * core/rawdb: better documentation	3 years ago
Martin Holst Swende	9b99e3dfe0	core/rawdb: fix datarace in freezer (#22728 ) The Append / truncate operations were racy. When a datafile reaches 2Gb, a new file is needed. For this operation, we require a writelock, which is not needed in the 99.99% of all cases where the data does fit in the current head-file. This transition from readlock to writelock was incorrect, and as the readlock was released, a truncate operation could slip in between, and truncate the data. This would have been fine, however, the Append operation continued writing as if no truncation had occurred, e.g writing item 5 where item 0 should reside. This PR changes the behaviour, so that if when we run into the situation that a new file is needed, it aborts, and retries, this time with a writelock. The outcome of the situation described above, running on this PR, would instead be that the Append operation exits with a failure.	4 years ago
Martin Holst Swende	271e5b7fc9	cmd/geth: add db-command to inspect freezer index (#22633 ) This PR makes it easier to inspect the freezer index, which could be useful to investigate things like #22111	4 years ago
Martin Holst Swende	681618275c	core: speed up header import (#21967 ) This PR implements the following modifications - Don't shortcut check if block is present, thus avoid disk lookup - Don't check hash ancestry in early-check (it's still done in parallel checker) - Don't check time.Now for every single header Charts and background info can be found here: https://github.com/holiman/headerimport/blob/main/README.md With these changes, writing 1M headers goes down to from 80s to 62s.	4 years ago
Péter Szilágyi	5655dce3b8	core/rawdb: only complain loudly if truncating many items	4 years ago
zhangsoledad	bcf19bc4be	core/rawdb: swap tailId and itemOffset for deleted items in freezer (#21220 ) * fix(freezer): tailId filenum offset were misplaced * core/rawdb: assume first item in freezer always start from zero	4 years ago
Marius van der Wijden	2a836bb259	core/rawdb: fix data race between Retrieve and Close (#20919 ) * core/rawdb: fixed data race between retrieve and close closes https://github.com/ethereum/go-ethereum/issues/20420 * core/rawdb: use non-atomic load while holding mutex	5 years ago
Péter Szilágyi	72d5a27a39	core, metrics, p2p: switch some invalid counters to gauges	5 years ago
Christian Muehlhaeuser	57fc1d21e1	cmd/geth, core/rawdb: add missing error checks (#19871 ) * Added missing error checks Add error handling where we assign err a value, but don't check for it being nil. * core/rawdb: tiny style nit	5 years ago
Péter Szilágyi	b02958b9c5	core, ethdb, metrics, p2p: expose various counter metrics for grafana	5 years ago
Frank Szendzielarz	f9c0e093ed	core/rawdb: avoid O_APPEND (#19676 ) * Fix file system access for Windows * Encapsulate file accesses * Style fixes	5 years ago
Péter Szilágyi	f35975ea21	core/rawdb, eth/downloader: align 64bit atomic fields	6 years ago
Péter Szilágyi	9eba3a9fff	cmd/geth, core/rawdb: seamless freezer consistency, friendly removedb	6 years ago
gary rong	37d280da41	core, cmd, vendor: fixes and database inspection tool (#15 ) * core, eth: some fixes for freezer * vendor, core/rawdb, cmd/geth: add db inspector * core, cmd/utils: check ancient store path forceily * cmd/geth, common, core/rawdb: a few fixes * cmd/geth: support windows file rename and fix rename error * core: support ancient plugin * core, cmd: streaming file copy * cmd, consensus, core, tests: keep genesis in leveldb * core: write txlookup during ancient init * core: bump database version	6 years ago
Martin Holst Swende	42c746d6f4	freezer: disable compression on hashes and difficulties (#14 ) * freezer: disable compression on hashes and difficulties * core/rawdb: address review concerns * core/rawdb: address review concerns	6 years ago
Martin Holst Swende	331de17e4d	core/rawdb: support starting offset for future deletion	6 years ago
gary rong	80469bea0c	all: integrate the freezer with fast sync * all: freezer style syncing core, eth, les, light: clean up freezer relative APIs core, eth, les, trie, ethdb, light: clean a bit core, eth, les, light: add unit tests core, light: rewrite setHead function core, eth: fix downloader unit tests core: add receipt chain insertion test core: use constant instead of hardcoding table name core: fix rollback core: fix setHead core/rawdb: remove canonical block first and then iterate side chain core/rawdb, ethdb: add hasAncient interface eth/downloader: calculate ancient limit via cht first core, eth, ethdb: lots of fixes * eth/downloader: print ancient disable log only for fast sync	6 years ago
Martin Holst Swende	b69bdc2a4f	freezer: implement split files for data * freezer: implement split files for data * freezer: add tests * freezer: close old head-file when opening next * freezer: fix truncation * freezer: more testing around close/open * rawdb/freezer: address review concerns * freezer: fix minor review concerns * freezer: fix remaining concerns + testcases around truncation * freezer: docs * freezer: implement multithreading * core/rawdb: fix freezer nitpicks + change offsets to uint32 * freezer: preopen files, simplify lock constructs * freezer: delete files during truncation	6 years ago
Péter Szilágyi	006c21efc7	cmd, core, eth, les, node: chain freezer on top of db rework	6 years ago

38 Commits (dac54e31a7456a364d07f238749c8354a110689b)