lighthouse

mirror of https://github.com/sigp/lighthouse.git synced 2026-06-01 13:47:16 +00:00

Author	SHA1	Message	Date
Michael Sproul	c7bb3b00e4	Fix lookups of the block at `oldest_block_slot` (#7693 ) Closes: - https://github.com/sigp/lighthouse/issues/7690 Another checkpoint sync related fix! See issue for a description of the bug. We fix it by just loading the block root of the `oldest_block_slot`, rather than trying to load the slot prior, which will always fail.	2025-07-02 23:40:04 +00:00
Jimmy Chen	b35854b71f	Record v2 beacon blocks http api metrics separately (#7692 ) This PR adds v2 beacon block paths to the function that records http api usage, so they don't just get recorded as "/v2/beacon" like below: <img width="934" alt="image" src="https://github.com/user-attachments/assets/8b669f0a-2821-46ee-a30a-0e344d3e63c1" />	2025-07-02 08:47:35 +00:00
Michael Sproul	a459a9af98	Fix and test checkpoint sync from genesis (#7689 ) Fix a bug involving checkpoint sync from genesis reported by Sunnyside labs. Ensure that the store's `anchor` is initialised prior to storing the genesis state. In the case of checkpoint sync from genesis, the genesis state will be in the _hot DB_, so we need the hot DB metadata to be initialised in order to store it. I've extended the existing checkpoint sync tests to cover this case as well. There are some subtleties around what the `state_upper_limit` should be set to in this case. I've opted to just enable state reconstruction from the start in the test so it gets set to 0, which results in an end state more consistent with the other test cases (full state reconstruction). This is required because we can't meaningfully do any state reconstruction when the split slot is 0 (there is no range of frozen slots to reconstruct).	2025-07-02 04:50:33 +00:00
Jimmy Chen	fcc602a787	Update fulu network configs and add `MIN_EPOCHS_FOR_DATA_COLUMN_SIDECARS_REQUESTS` (#7646 ) - #6240 - Bring built-in network configs up to date with latest consensus-spec PeerDAS configs. - Add `MIN_EPOCHS_FOR_DATA_COLUMN_SIDECARS_REQUESTS` and use it to determine data availability window after the Fulu fork.	2025-07-02 02:38:25 +00:00
Pawan Dhananjay	69c9c7038a	Use prepare_beacon_proposer endpoint for validator custody registration (#7681 ) N/A This PR switches to using `prepare_beacon_proposer` instead of `beacon_committee_subscriptions` endpoint to register validators with the custody context. We currently use the `beacon_committee_subscriptions` endpoint for registering validators in the custody context. Using the subscriptions endpoint has a few disadvantages: 1. The lighthouse VC tries to optimise the number of calls it makes to this endpoint to reduce the load on the subscriptions endpoint. So we would be getting different a subset of the total number of validators in each call. This will lead to a ramp up of the validator custody units instead of a one time bump. For e.g. see these logs ``` Jun 30 22:36:05.012 DEBUG Validator count at head updated old_count: 0, new_count: 19 Jun 30 22:36:11.016 DEBUG Validator count at head updated old_count: 19, new_count: 24 Jun 30 22:36:17.017 DEBUG Validator count at head updated old_count: 24, new_count: 27 Jun 30 22:36:23.020 DEBUG Validator count at head updated old_count: 27, new_count: 32 Jun 30 22:36:29.016 DEBUG Validator count at head updated old_count: 32, new_count: 36 Jun 30 22:36:35.005 DEBUG Validator count at head updated old_count: 36, new_count: 42 Jun 30 22:36:41.014 DEBUG Validator count at head updated old_count: 42, new_count: 44 Jun 30 22:36:47.017 DEBUG Validator count at head updated old_count: 44, new_count: 46 Jun 30 22:36:53.007 DEBUG Validator count at head updated old_count: 46, new_count: 48 Jun 30 22:36:59.009 DEBUG Validator count at head updated old_count: 48, new_count: 49 Jun 30 22:37:05.014 DEBUG Validator count at head updated old_count: 49, new_count: 50 Jun 30 22:37:11.007 DEBUG Validator count at head updated old_count: 50, new_count: 53 Jun 30 22:37:17.007 DEBUG Validator count at head updated old_count: 53, new_count: 55 Jun 30 22:37:35.008 DEBUG Validator count at head updated old_count: 55, new_count: 58 Jun 30 22:37:41.007 DEBUG Validator count at head updated old_count: 58, new_count: 59 Jun 30 22:37:53.010 DEBUG Validator count at head updated old_count: 59, new_count: 60 Jun 30 22:38:05.013 DEBUG Validator count at head updated old_count: 60, new_count: 61 Jun 30 22:38:23.006 DEBUG Validator count at head updated old_count: 61, new_count: 62 Jun 30 22:38:29.009 DEBUG Validator count at head updated old_count: 62, new_count: 63 Jun 30 22:38:41.009 DEBUG Validator count at head updated old_count: 63, new_count: 64 ``` 2. Different VCs would probably have different behaviours in terms of sending subscriptions In contrast, the `prepare_beacon_proposer` endpoint usage would be more standard across different VCs without any filtering of validators. Not doing so could mean potentially missing proposals so VCs are incentivised to make this call on any change in the validators managed by them. Lighthouse calls this endpoint every slot.	2025-07-02 02:38:22 +00:00
Jimmy Chen	41742ce2bd	Update `SAMPLES_PER_SLOT` to be number of custody groups instead of data columns (#7683 ) Update `SAMPLES_PER_SLOT` to be number of custody groups instead of data columns. This should have no impact on the current implementation as config currently maintains a `group:subnet:column` ratio of `1:1:1`. In short, this PR doesn't change anything for Fusaka, but ensures compliance with the spec and potential future changes. I've added separate methods to compute sampling columns and custody groups for clarity: `spec.sampling_size_columns` and `spec.sampling_size_custod_groups` See the clarifications in this PR for more details: https://github.com/ethereum/consensus-specs/pull/4251	2025-07-02 00:08:40 +00:00
Pawan Dhananjay	e305cb1b92	Custody persist fix (#7661 ) N/A Persist the epoch -> cgc values. This is to ensure that `ValidatorRegistrations::latest_validator_custody_requirement` always returns a `Some` value post restart assuming the `epoch_validator_custody_requirements` map has been updated in the previous runs.	2025-07-01 06:06:37 +00:00
chonghe	257d270718	Add voluntary exit via validator manager (#6612 ) * #4303 * #4804 -Add voluntary exit feature to the validator manager -Add delete all validators by using the keyword "all"	2025-07-01 03:07:49 +00:00
Michael Sproul	c1f94d9b7b	Test database schema stability (#7669 ) This PR implements some heuristics to check for breaking database changes. The goal is to prevent accidental changes to the database schema occurring without a version bump.	2025-06-30 10:55:47 +00:00
Michael Sproul	25ea8a83b7	Add Michael as codeowner for store crate (#7667 ) I'm adding myself as a codeowner for the `store` crate so that I can more easily keep track of database-related PRs.	2025-06-30 10:47:52 +00:00
Jimmy Chen	e45ba846ae	Increase http client default timeout to 2s in `http-api` tests. (#7673 ) The `sync_contributions_across_fork_with_skip_slots` test have been quite flaky recently on CI, we suspect it might be caused by the recent introduction of a `default` timeout in #7400, and CI is failing to consistently process those http requests within 1s likely due to limited resources. This PR increases the `default` timeout to 2s in tests to avoid flaky tests, but keeps the remaining timeout the same (1s). https://github.com/sigp/lighthouse/actions/runs/15965113170/job/45023976021 ``` FAIL [ 8.945s] http_api::bn_http_api_tests fork_tests::sync_contributions_across_fork_with_skip_slots stdout ─── running 1 test test fork_tests::sync_contributions_across_fork_with_skip_slots ... FAILED failures: failures: fork_tests::sync_contributions_across_fork_with_skip_slots test result: FAILED. 0 passed; 1 failed; 0 ignored; 0 measured; 175 filtered out; finished in 8.91s stderr ─── thread 'fork_tests::sync_contributions_across_fork_with_skip_slots' panicked at beacon_node/http_api/tests/fork_tests.rs:239:10: called `Result::unwrap()` on an `Err` value: HttpClient(url: http://127.0.0.1:41793/, kind: timeout, detail: operation timed out) note: run with `RUST_BACKTRACE=1` environment variable to display a backtrace ```	2025-06-30 10:47:49 +00:00
Michael Sproul	6be646ca11	Bump DB schema to v25 (#7666 ) When we removed the eth1 data, I wrote a v25 schema upgrade to delete the data on disk: - https://github.com/sigp/lighthouse/pull/7133 However, I forgot to update the current schema version, so this change was never actioned. This PR updates the current schema version to v25 so that the migration runs.	2025-06-30 05:52:28 +00:00
Nikola Ristić	2d759f78be	Fix beacon_chain metrics descriptions (#6576 ) n/a Just fixed a small mistake in the metrics description.	2025-06-30 00:11:14 +00:00
Odinson	6ea5f14b39	feat: better error message for light_client/bootstrap endpoint (#7597 ) Fixes #7229	2025-06-28 03:55:39 +00:00
Jimmy Chen	522e00f48d	Fix incorrect `waker` update condition (#7656 ) This bug was first found and partially fixed by @VolodymyrBg in #7317 - this PR applies the same fix everywhere else. The old logic updated the waker when it already matched the context, and did nothing when it was stale: ```rust if waker.will_wake(cx.waker()) { self.waker = Some(cx.waker().clone()); } ``` This is the wrong way around. We only want to update the waker if it doesn't match the current context: ```rust if !waker.will_wake(cx.waker()) { self.waker = Some(cx.waker().clone()); } ``` I don't think we've ever noticed any issues, but it’s a subtle bug that could lead to missed wakeups.	2025-06-27 19:01:46 +00:00
Jimmy Chen	83cad25d98	Fix Rust 1.88 clippy errors & execution engine tests (#7657 ) Fix Rust 1.88 clippy errors.	2025-06-27 18:21:17 +00:00
Pawan Dhananjay	9b1f3ed9d1	Add gossip check (#7652 ) N/A Add an additional gossip condition.	2025-06-27 00:26:38 +00:00
Daniel Knopik	a0a6b9300f	Do not compute sync selection proofs for the sync duty at the current slot (#7551 )	2025-06-25 06:22:24 +00:00
chonghe	8e3c5d1524	Rust 1.89 compiler lint fix (#7644 ) Fix lints for Rust 1.89 beta compiler	2025-06-25 05:33:17 +00:00
Eitan Seri-Levi	56b2d4b525	Remove instrumenting log level (#7636 ) https://github.com/sigp/lighthouse/issues/7155 Theres some additional places we set instrumenting log levels that wasn't covered in #7620	2025-06-24 06:29:10 +00:00
Michael Sproul	fd643c310c	Un-ignore EF test for v1.6.0-alpha.1 (#7632 ) Closes: - https://github.com/sigp/lighthouse/issues/7547 Run the test that was previously ignored when we were between spec versions.	2025-06-23 13:11:46 +00:00
chonghe	cef04ee2ee	Implement `validator_identities` Beacon API endpoint (#7462 ) * #7442	2025-06-23 08:37:49 +00:00
Pawan Dhananjay	3fefda68e5	Send byrange responses in the correct requested range (#7611 ) N/A For responding to by_range requests , we should ideally only respond with items in the range `req.start_slot()..req.start_slot() + req.count`. We were not filtering the generated response for blobs and data columns, only for blocks. This PR adds the filtering for the sidecars as well.	2025-06-23 06:57:37 +00:00
Mac L	e34a9a0c65	Allow the `--beacon-nodes` list to be updated at runtime (#6551 ) Adds a new `/lighthouse` API call to the VC which allows the list of beacon nodes to be updated dynamically at runtime. An entirely new beacon node list is provided to the VC so it effectively adds, removes or reorders nodes to match the new list. This can then be used in Siren, which will enable a "drag to reorder" system along with adding and removing beacon nodes while the VC is on. This will make it unnecessary to reboot the VC when users want to simply add or remove a BN from the list.	2025-06-23 03:59:34 +00:00
Pawan Dhananjay	11bcccb353	Remove all prod eth1 related code (#7133 ) N/A After the electra fork which includes EIP 6110, the beacon node no longer needs the eth1 bridging mechanism to include new deposits as they are provided by the EL as a `deposit_request`. So after electra + a transition period where the finalized bridge deposits pre-fork are included through the old mechanism, we no longer need the elaborate machinery we had to get deposit contract data from the execution layer. Since holesky has already forked to electra and completed the transition period, this PR basically checks to see if removing all the eth1 related logic leads to any surprises.	2025-06-23 03:00:07 +00:00
Age Manning	d50924677a	Remove instrumenting log level (#7620 ) I think this should resolve #7155 This removes the level field from the instrumenting we were doing across a range of functions. The level will now default to the level of the log.	2025-06-20 07:44:59 +00:00
Eitan Seri-Levi	f67084a571	Remove reprocess channel (#7437 ) Partially https://github.com/sigp/lighthouse/issues/6291 This PR removes the reprocess event channel from being externally exposed. All work events are now sent through the single `BeaconProcessorSend` channel. I've introduced a new `Work::Reprocess` enum variant which we then use to schedule jobs for reprocess. I've also created a new scheduler module which will eventually house the different scheduler impls. This is all needed as an initial step to generalize the beacon processor A "full" implementation for the generalized beacon processor can be found here https://github.com/sigp/lighthouse/pull/6448 I'm going to try to break up the full implementation into smaller PR's so it can actually be reviewed	2025-06-20 02:52:16 +00:00
Lion - dapplion	dd98534158	Hierarchical state diffs in hot DB (#6750 ) This PR implements https://github.com/sigp/lighthouse/pull/5978 (tree-states) but on the hot DB. It allows Lighthouse to massively reduce its disk footprint during non-finality and overall I/O in all cases. Closes https://github.com/sigp/lighthouse/issues/6580 Conga into https://github.com/sigp/lighthouse/pull/6744 ### TODOs - [x] Fix OOM in CI https://github.com/sigp/lighthouse/pull/7176 - [x] optimise store_hot_state to avoid storing a duplicate state if the summary already exists (should be safe from races now that pruning is cleaner) - [x] mispelled: get_ancenstor_state_root - [x] get_ancestor_state_root should use state summaries - [x] Prevent split from changing during ancestor calc - [x] Use same hierarchy for hot and cold ### TODO Good optimization for future PRs - [ ] On the migration, if the latest hot snapshot is aligned with the cold snapshot migrate the diffs instead of the full states. ``` align slot time 10485760 Nov-26-2024 12582912 Sep-14-2025 14680064 Jul-02-2026 ``` ### TODO Maybe things good to have - [ ] Rename anchor_slot https://github.com/sigp/lighthouse/compare/tree-states-hot-rebase-oom...dapplion:lighthouse:tree-states-hot-anchor-slot-rename?expand=1 - [ ] Make anchor fields not public such that they must be mutated through a method. To prevent un-wanted changes of the anchor_slot ### NOTTODO - [ ] Use fork-choice and a new method [`descendants_of_checkpoint`](`ca2388e196 (diff-046fbdb517ca16b80e4464c2c824cf001a74a0a94ac0065e635768ac391062a8)`) to filter only the state summaries that descend of finalized checkpoint]	2025-06-19 02:43:25 +00:00
Eitan Seri-Levi	6786b9d12a	Single attestation "Full" implementation (#7444 ) #6970 This allows for us to receive `SingleAttestation` over gossip and process it without converting. There is still a conversion to `Attestation` as a final step in the attestation verification process, but by then the `SingleAttestation` is fully verified. I've also fully removed the `submitPoolAttestationsV1` endpoint as its been deprecated I've also pre-emptively deprecated supporting `Attestation` in `submitPoolAttestationsV2` endpoint. See here for more info: https://github.com/ethereum/beacon-APIs/pull/531 I tried to the minimize the diff here by only making the "required" changes. There are some unnecessary complexities with the way we manage the different attestation verification wrapper types. We could probably consolidate this to one wrapper type and refactor this even further. We could leave that to a separate PR if we feel like cleaning things up in the future. Note that I've also updated the test harness to always submit `SingleAttestation` regardless of fork variant. I don't see a problem in that approach and it allows us to delete more code :)	2025-06-17 09:01:26 +00:00
Jimmy Chen	3d2d65bf8d	Advertise `--advertise-false-custody-group-count` for testing PeerDAS (#7593 ) #6973	2025-06-16 11:10:28 +00:00
Jimmy Chen	6135f417a2	Add data columns sidecars debug beacon API (#7591 ) Beacon API spec PR: https://github.com/ethereum/beacon-APIs/pull/537	2025-06-15 14:20:16 +00:00
diegomrsantos	4fc0665ccd	Add more context to Late Block Re-orgs (#7592 ) Giving more context about late block re-orgs would make the concept easier to grasp for newcomers. Add more context to this section in the Lighthouse Book.	2025-06-14 03:45:35 +00:00
Akihito Nakano	dc5f5af3eb	Fix flaky test_rpc_block_reprocessing (#7595 ) The test occasionally fails, likely because the 10ms fixed delay after block processing isn't insufficient when the system is under load. https://github.com/sigp/lighthouse/pull/7522#issuecomment-2914595667 Replace single assertion with retry loop.	2025-06-14 00:54:19 +00:00
Daniel Knopik	ccd99c138c	Wait before column reconstruction (#7588 )	2025-06-13 18:19:06 +00:00
Jimmy Chen	a65f78222d	Drop stale registrations without reducing CGC (#7594 ) Currently the validator effective balance used for computing PeerDAS custody group count is only updated when the validator subscribes to the BN via `validator/beacon_committee_subscriptions`. If a validator stops registering with the node, the effective balance gets outdated and stays in the BN memory until the next restart. They are no longer required for CGC computation, as long as the CGC never reduces as per the spec, therefore they can be dropped.	2025-06-13 14:30:43 +00:00
Daniel Knopik	5472cb8500	Batch verify KZG proofs for getBlobsV2 (#7582 )	2025-06-12 14:35:14 +00:00
Pawan Dhananjay	9803d69d80	Implement status v2 version (#7590 ) N/A Implements status v2 as defined in https://github.com/ethereum/consensus-specs/pull/4374/	2025-06-12 07:17:06 +00:00
Pawan Dhananjay	5f208bb858	Implement basic validator custody framework (no backfill) (#7578 ) Resolves #6767 This PR implements a basic version of validator custody. - It introduces a new `CustodyContext` object which contains info regarding number of validators attached to a node and the custody count they contribute to the cgc. - The `CustodyContext` is added in the da_checker and has methods for returning the current cgc and the number of columns to sample at head. Note that the logic for returning the cgc existed previously in the network globals. - To estimate the number of validators attached, we use the `beacon_committee_subscriptions` endpoint. This might overestimate the number of validators actually publishing attestations from the node in the case of multi BN setups. We could also potentially use the `publish_attestations` endpoint to get a more conservative estimate at a later point. - Anytime there's a change in the `custody_group_count` due to addition/removal of validators, the custody context should send an event on a broadcast channnel. The only subscriber for the channel exists in the network service which simply subscribes to more subnets. There can be additional subscribers in sync that will start a backfill once the cgc changes. TODO - [ ] NOT REQUIRED: Currently, the logic only handles an increase in validator count and does not handle a decrease. We should ideally unsubscribe from subnets when the cgc has decreased. - [ ] NOT REQUIRED: Add a service in the `CustodyContext` that emits an event once `MIN_EPOCHS_FOR_BLOB_SIDECARS_REQUESTS ` passes after updating the current cgc. This event should be picked up by a subscriber which updates the enr and metadata. - [x] Add more tests	2025-06-11 18:10:06 +00:00
Pawan Dhananjay	076a1c3fae	Data column sidecar event (#7587 ) N/A Implement events for data column sidecar https://github.com/ethereum/beacon-APIs/pull/535	2025-06-11 16:39:22 +00:00
chonghe	7416d06dce	Add genesis sync test to CI (#7561 ) * #7550 Use existing code from @jimmygchen in #7530 and modify for genesis sync test. Thanks @jimmygchen !	2025-06-11 09:51:37 +00:00
Jimmy Chen	8c6abc0b69	Optimise parallelism in compute cells operations by zipping first (#7574 ) We're seeing slow KZG performance on `fusaka-devnet-0` and looking for optimisations to improve performance. Zipping the list first then `into_par_iter` shows a 10% improvement in performance benchmark, i suspect this might be even more material when running on a beacon node. Before: ``` blobs_to_data_column_sidecars_20 time: [11.583 ms 12.041 ms 12.534 ms] Found 5 outliers among 100 measurements (5.00%) ``` After: ``` blobs_to_data_column_sidecars_20 time: [10.506 ms 10.724 ms 10.982 ms] change: [-14.925% -10.941% -6.5452%] (p = 0.00 < 0.05) Performance has improved. Found 6 outliers among 100 measurements (6.00%) ```	2025-06-09 12:41:14 +00:00
ethDreamer	b08d49c4cb	Changes for `fusaka-devnet-1` (#7559 ) Changes for [fusaka-devnet-1](https://notes.ethereum.org/@ethpandaops/fusaka-devnet-1) [Consensus Specs v1.6.0-alpha.1](https://github.com/ethereum/consensus-specs/pull/4346) * [EIP-7917: Deterministic Proposer Lookahead](https://eips.ethereum.org/EIPS/eip-7917) * [EIP-7892: Blob Parameter Only Hardforks](https://eips.ethereum.org/EIPS/eip-7892)	2025-06-09 09:10:08 +00:00
Akihito Nakano	170cd0f587	Store the libp2p/discv5 logs when stopping local-testnet (#7579 ) The libp2p/discv5 logs are not stored when stopping local-testnet. Store the `beacon/logs` directory to [Kurtosis Files Artifacts](https://docs.kurtosis.com/advanced-concepts/files-artifacts/) so that they are downloaded locally by `kurtosis enclave dump`.	2025-06-08 03:21:41 +00:00
Jimmy Chen	b2e8b67e34	Reduce number of basic sim test nodes from 7 to 4 (#7566 ) Our basic sim test has been [flaky](https://github.com/sigp/lighthouse/actions/runs/15458818777/job/43515966229) for some time, and seems like it has gotten worse since electra fork was added to it in #7199. It looks like the github runner is struggling with the load, currently it runs 7 nodes on a 4 CPU runner, which is definitely too much. We could consider moving this to run on our self hosted runner - but I think running 7 nodes is unnecessary and we can probably trim test this down. Reduce number of basic sim test nodes from 7 (3 BN + 3 Proposer BN + 1 extra) to 4 (2 BN + 1 Proposer BN + 1 extra). If we want to run more nodes, we'd have to consider running on self hosted runners.	2025-06-06 03:51:51 +00:00
Jimmy Chen	e098f66738	Update kurtosis config and EL images (#7570 ) Update kurtosis config to start from electra genesis. https://github.com/sigp/lighthouse/issues/6826#issuecomment-2900375344	2025-06-05 16:20:33 +00:00
Justin Traglia	2f807e21be	Add support for nightly tests (#7538 ) This PR adds the ability to download [nightly reference tests from the consensus-specs repo](https://github.com/ethereum/consensus-specs/actions/workflows/generate_vectors.yml). This will be used by spec maintainers to ensure that there are no unexpected test failures prior to new releases. Also, we will keep track of test compliance with [this website](https://jtraglia.github.io/nyx/); eventually this will be integrated into Hive. * A new script (`download_test_vectors.sh`) is added to handle downloads. * The logic for downloading GitHub artifacts is a bit complex. * Rename the variables which store test versions: * `TESTS_TAG` to `CONSENSUS_SPECS_TEST_VERSION`. * `BLS_TEST_TAG` to `BLS_TEST_VERSION`, for consistency. * Delete tarballs after extracting them. * I see no need to keep these; they just use extra disk. * Consolidate `clean` rules into a single rule. * Do `clean` prior to downloading/extracting tests. * Remove `CURL` variable with GitHub token; don't need it for downloading releases. * Do `mkdir -p` when creating directories. * Probably more small stuff...	2025-06-05 12:28:06 +00:00
Lion - dapplion	d457ceeaaf	Don't create child lookup if parent is faulty (#7118 ) Issue discovered on PeerDAS devnet (node `lighthouse-geth-2.peerdas-devnet-5.ethpandaops.io`). Summary: - A lookup is created for block root `0x28299de15843970c8ea4f95f11f07f75e76a690f9a8af31d354c38505eebbe12` - That block or a parent is faulty and `0x28299de15843970c8ea4f95f11f07f75e76a690f9a8af31d354c38505eebbe12` is added to the failed chains cache - We later receive a block that is a child of a child of `0x28299de15843970c8ea4f95f11f07f75e76a690f9a8af31d354c38505eebbe12` - We create a lookup, which attempts to process the child of `0x28299de15843970c8ea4f95f11f07f75e76a690f9a8af31d354c38505eebbe12` and hit a processor error `UnknownParent`, hitting this line `bf955c7543/beacon_node/network/src/sync/block_lookups/mod.rs (L686-L688)` `search_parent_of_child` does not create a parent lookup because the parent root is in the failed chain cache. However, we have already marked the child as awaiting the parent. This results in an inconsistent state of lookup sync, as there's a lookup awaiting a parent that doesn't exist. Now we have a lookup (the child of `0x28299de15843970c8ea4f95f11f07f75e76a690f9a8af31d354c38505eebbe12`) that is awaiting a parent lookup that doesn't exist: hence stuck. ### Impact This bug can affect Mainnet as well as PeerDAS devnets. This bug may stall lookup sync for a few minutes (up to `LOOKUP_MAX_DURATION_STUCK_SECS = 15 min`) until the stuck prune routine deletes it. By that time the root will be cleared from the failed chain cache and sync should succeed. During that time the user will see a lot of `WARN` logs when attempting to add each peer to the inconsistent lookup. We may also sync the block through range sync if we fall behind by more than 2 epochs. We may also create the parent lookup successfully after the failed cache clears and complete the child lookup. This bug is triggered if: - We have a lookup that fails and its root is added to the failed chain cache (much more likely to happen in PeerDAS networks) - We receive a block that builds on a child of the block added to the failed chain cache Ensure that we never create (or leave existing) a lookup that references a non-existing parent. I added `must_use` lints to the functions that create lookups. To fix the specific bug we must recursively drop the child lookup if the parent is not created. So if `search_parent_of_child` returns `false` now return `LookupRequestError::Failed` instead of `LookupResult::Pending`. As a bonus I have a added more logging and reason strings to the errors	2025-06-05 08:53:43 +00:00
Jimmy Chen	9a4972053e	Add e2e sync tests to CI (#7530 ) This PR adds the following sync tests to CI workflow - triggered when a PR is labeled `syncing` - to ensure we have some e2e coverage on basic sync scenarios: - [x] checkpoint sync to a live network (covers range and backfill sync for _current_ fork) - [x] checkpoint sync to a running devnet (covers range and backfill sync for _next_ fork) It seems to work fine running on github hosted runners - but if performance become an issue we could switch to using self hosted runners for sepolia sync test. (standard CPU runners have 4 CPU, 16 GB ram - i think it _should_ be enough on sepolia / devnet networks) The following tests have been removed from this PR and moved to a separate issue *(#7550) - [x] genesis sync on a local devnet (covers current and next fork) - [x] brief shutdown and restart (covers lookup sync) - [x] longer shutdown and restart (covers range sync) I'm hoping to keep these e2e test maintenance effort to a minimum - hopefully longer term we could have some generic e2e tests that works for all clients and the maintenance effort can be spread across teams. ### Latest test run: https://github.com/sigp/lighthouse/actions/runs/15411744248 ### Results: <img width="687" alt="image" src="https://github.com/user-attachments/assets/c7178291-7b39-4f3b-a339-d3715eb16081" /> <img width="693" alt="image" src="https://github.com/user-attachments/assets/a8fc3520-296c-4baf-ae1e-1e887e660a3c" /> #### logs are available as artifacts: <img width="629" alt="image" src="https://github.com/user-attachments/assets/3c0e1cd7-9c94-4d0c-be62-5e45179ab8f3" />	2025-06-05 08:31:55 +00:00
chonghe	dcee76c0dc	Update key generation in validator manager (#7548 ) #7518 breaks the key generation process in `validator_manager/test_vectors/generate.py` after updating to `ethstaker-deposit-cli`. This PR updates the key generation process, tested and successfully generated the deposit data JSON files.	2025-06-05 06:34:33 +00:00
ethDreamer	2d9fc34d43	Fulu EF tests v1.6.0-alpha.0 (#7540 ) Update to EF tests v1.6.0-alpha.0	2025-06-04 06:34:12 +00:00

1 2 3 4 5 ...

6912 Commits