Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Add copy buttons to all
 blocks\n(function() {\n function addCopyButtons() {\n document.querySelectorAll('pre code').forEach(function(codeBlock) {\n if (codeBlock.parentElement.hasAttribute('data-copy-added')) return;\n codeBlock.parentElement.setAttribute('data-copy-added', 'true');\n \n var btn = document.createElement('button');\n btn.textContent = 'Copy';\n btn.style.cssText = 'position:absolute;top:4px;right:4px;padding:2px 8px;font-size:11px;background:#4ecdc4;border:none;border-radius:4px;color:#1a1a2e;cursor:pointer;opacity:0.7;transition:opacity 0.2s;';\n btn.onmouseover = function() { this.style.opacity = '1'; };\n btn.onmouseout = function() { this.style.opacity = '0.7'; };\n btn.onclick = function() {\n navigator.clipboard.writeText(codeBlock.textContent).then(function() {\n btn.textContent = 'Copied!';\n setTimeout(function() { btn.textContent = 'Copy'; }, 1500);\n });\n };\n codeBlock.parentElement.style.position = 'relative';\n codeBlock.parentElement.appendChild(btn);\n });\n }\n \n addCopyButtons();\n \n // Re-run on dynamic content\n var observer = new MutationObserver(addCopyButtons);\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Add Copy Buttons to Code Blocks");
}
} catch(__e) { console.warn('[Userscript:Add Copy Buttons to Code Blocks]', __e); }
})();
(function(){
try {
var __m = "github.com";
var __re = new RegExp('^' + "github\\.com" + '
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Force GitHub README to respect dark mode\n(function() {\n var style = document.createElement('style');\n style.textContent = '\n .markdown-body {\n color-scheme: dark light;\n }\n .markdown-body pre { background: #161b22 !important; }\n .markdown-body code { background: rgba(110, 118, 129, 0.4) !important; }\n .markdown-body table th, .markdown-body table td { border-color: #30363d !important; }\n .markdown-body img { background: #0d1117; }\n .markdown-body blockquote { border-left-color: #8b949e; }\n .markdown-body hr { border-color: #30363d; }\n ';\n document.head.appendChild(style);\n})();", "GitHub Dark Mode README Fix"); } } catch(__e) { console.warn('[Userscript:GitHub Dark Mode README Fix]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Highlight search terms from Google/DuckDuckGo/Bing referrer\n(function() {\n var ref = document.referrer;\n var terms = [];\n \n if (ref.includes('google.com') || ref.includes('duckduckgo.com') || ref.includes('bing.com')) {\n var url = new URL(ref);\n var q = url.searchParams.get('q') || url.searchParams.get('p');\n if (q) {\n terms = q.split(/\\s+/).filter(function(t) { return t.length > 2; });\n }\n }\n \n if (terms.length === 0) return;\n \n var style = document.createElement('style');\n style.textContent = '.userscript-highlight { background: #fbbf24; color: #1a1a2e; padding: 1px 3px; border-radius: 2px; }';\n document.head.appendChild(style);\n \n function highlight(node) {\n if (node.nodeType === 3) { // text node\n var text = node.textContent;\n var found = false;\n terms.forEach(function(term) {\n var regex = new RegExp('(' + term.replace(/[.*+?^${}()|[\\]\\\\]/g, '\\\\') + ')', 'gi');\n if (regex.test(text)) {\n found = true;\n var frag = document.createDocumentFragment();\n var parts = text.split(regex);\n parts.forEach(function(part, i) {\n if (i % 2 === 0) {\n frag.appendChild(document.createTextNode(part));\n } else {\n var span = document.createElement('span');\n span.className = 'userscript-highlight';\n span.textContent = part;\n frag.appendChild(span);\n }\n });\n node.parentNode.replaceChild(frag, node);\n }\n });\n } else if (node.nodeType === 1 && node.childNodes) { // element\n var skipTags = ['SCRIPT', 'STYLE', 'NOSCRIPT', 'TEXTAREA', 'INPUT', 'SELECT'];\n if (!skipTags.includes(node.tagName)) {\n Array.from(node.childNodes).forEach(highlight);\n }\n }\n }\n \n highlight(document.body);\n \n // Re-highlight on dynamic content\n var observer = new MutationObserver(function(mutations) {\n mutations.forEach(function(m) {\n m.addedNodes.forEach(function(node) {\n if (node.nodeType === 1 || node.nodeType === 3) highlight(node);\n });\n });\n });\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Highlight Search Terms"); } } catch(__e) { console.warn('[Userscript:Highlight Search Terms]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Strip utm_, fbclid, gclid, etc. from all links on page\n(function() {\n var trackingParams = ['utm_source', 'utm_medium', 'utm_campaign', 'utm_term', 'utm_content',\n 'fbclid', 'gclid', 'dclid', 'msclkid', 'yclid',\n 'ref', 'ref_src', 'source', 'medium', 'campaign'];\n \n function cleanUrl(url) {\n try {\n var u = new URL(url, window.location.origin);\n var changed = false;\n trackingParams.forEach(function(p) {\n if (u.searchParams.has(p)) {\n u.searchParams.delete(p);\n changed = true;\n }\n });\n return changed ? u.toString() : url;\n } catch (e) {\n return url;\n }\n }\n \n function cleanLinks() {\n document.querySelectorAll('a[href]').forEach(function(a) {\n var clean = cleanUrl(a.href);\n if (clean !== a.href) a.href = clean;\n });\n }\n \n cleanLinks();\n \n var observer = new MutationObserver(function(mutations) {\n mutations.forEach(function(m) {\n m.addedNodes.forEach(function(node) {\n if (node.nodeType === 1) {\n if (node.tagName === 'A') cleanLinks();\n node.querySelectorAll('a[href]').forEach(function(a) {\n var clean = cleanUrl(a.href);\n if (clean !== a.href) a.href = clean;\n });\n }\n });\n });\n });\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Remove Tracking Parameters from Links"); } } catch(__e) { console.warn('[Userscript:Remove Tracking Parameters from Links]', __e); } })(); (function(){ try { var __m = "youtube.com"; var __re = new RegExp('^' + "youtube\\.com" + '
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Auto-enable theater mode on YouTube\n(function() {\n function tryTheater() {\n var btn = document.querySelector('button[aria-label=\"Theater mode\"], ytd-player #player button[title=\"Theater mode\"]');\n if (btn && !btn.classList.contains('activated')) {\n btn.click();\n }\n }\n \n // Try immediately\n tryTheater();\n \n // Try after navigation (SPA)\n var lastUrl = location.href;\n setInterval(function() {\n if (location.href !== lastUrl) {\n lastUrl = location.href;\n setTimeout(tryTheater, 500);\n }\n }, 1000);\n \n // Also try on player load\n var observer = new MutationObserver(tryTheater);\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "YouTube Theater Mode Default"); } } catch(__e) { console.warn('[Userscript:YouTube Theater Mode Default]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Remove or un-stick sticky/fixed headers that block content\n(function() {\n function unstick() {\n document.querySelectorAll('header, nav, [role=\"banner\"], .header, .navbar, .sticky, .fixed-top, [style*=\"position: fixed\"], [style*=\"position:sticky\"]').forEach(function(el) {\n if (el.style.position === 'fixed' || el.style.position === 'sticky' || \n getComputedStyle(el).position === 'fixed' || getComputedStyle(el).position === 'sticky') {\n el.style.position = 'static';\n el.style.top = 'auto';\n el.style.zIndex = 'auto';\n }\n });\n }\n \n unstick();\n \n var observer = new MutationObserver(unstick);\n observer.observe(document.body, { childList: true, subtree: true, attributes: true, attributeFilter: ['style', 'class'] });\n})();", "Kill Sticky Headers"); } } catch(__e) { console.warn('[Userscript:Kill Sticky Headers]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Universal Dark Mode - works on any site\n(function() {\n var enabled = true;\n \n function applyDarkMode() {\n if (!enabled) return;\n \n // Create style element if it doesn't exist\n var style = document.getElementById('universal-dark-mode-style');\n if (!style) {\n style = document.createElement('style');\n style.id = 'universal-dark-mode-style';\n document.head.appendChild(style);\n }\n \n // Dark mode CSS - inverts colors but preserves images/video\n style.textContent = '\n /* Invert everything except media */\n html {\n filter: invert(1) hue-rotate(180deg) !important;\n background: #1a1a2e !important;\n }\n \n /* Restore images, videos, iframes, canvas */\n img, video, iframe, canvas, svg, picture, [style*=\"background-image\"] {\n filter: invert(1) hue-rotate(180deg) !important;\n }\n \n /* Preserve specific elements that should not be inverted */\n .no-dark-mode, .no-dark-mode *,\n [data-theme=\"light\"], [data-theme=\"light\"],\n .ace_editor, .ace_editor *,\n .CodeMirror, .CodeMirror *,\n .monaco-editor, .monaco-editor *,\n .markdown-body pre, .markdown-body pre *,\n .highlight, .highlight *,\n pre code, pre code * {\n filter: none !important;\n }\n \n /* Fix common UI elements */\n .modal, .popup, .dropdown-menu, .tooltip, .popover {\n filter: invert(1) hue-rotate(180deg) !important;\n background: #2d2d44 !important;\n border-color: #444 !important;\n }\n \n /* Scrollbars */\n ::-webkit-scrollbar { background: #1a1a2e !important; }\n ::-webkit-scrollbar-thumb { background: #444 !important; }\n ::-webkit-scrollbar-thumb:hover { background: #555 !important; }\n \n /* Selection */\n ::selection { background: #4ecdc4 !important; color: #1a1a2e !important; }\n ::-moz-selection { background: #4ecdc4 !important; color: #1a1a2e !important; }\n ';\n }\n \n function removeDarkMode() {\n var style = document.getElementById('universal-dark-mode-style');\n if (style) style.remove();\n }\n \n // Toggle with Alt+Shift+D\n document.addEventListener('keydown', function(e) {\n if (e.altKey && e.shiftKey && e.key === 'D') {\n e.preventDefault();\n enabled = !enabled;\n if (enabled) {\n applyDarkMode();\n console.log('[Universal Dark Mode] Enabled');\n } else {\n removeDarkMode();\n console.log('[Universal Dark Mode] Disabled');\n }\n }\n });\n \n // Apply on load\n applyDarkMode();\n \n // Re-apply on dynamic content\n var observer = new MutationObserver(function(mutations) {\n if (enabled && !document.getElementById('universal-dark-mode-style')) {\n applyDarkMode();\n }\n });\n observer.observe(document.head, { childList: true });\n \n console.log('[Universal Dark Mode] Loaded - Press Alt+Shift+D to toggle');\n})();", "Universal Dark Mode"); } } catch(__e) { console.warn('[Userscript:Universal Dark Mode]', __e); } })(); })();
Skip to content

Locked splice bug fixes during channel reestablishment - #4624

Merged
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes
May 21, 2026
Merged

Locked splice bug fixes during channel reestablishment#4624
jkczyz merged 2 commits into
lightningdevkit:mainfrom
wpaulino:splice-locked-reestablish-fixes

Conversation

@wpaulino

Copy link
Copy Markdown
Contributor

This PR includes fixes for two bugs/edge cases when handling a locked splice during reestablishment. These were found by the chanmon_consistency fuzz target.

@wpaulinowpaulino added this to the 0.3 milestone May 19, 2026
@wpaulinowpaulino self-assigned this May 19, 2026
@ldk-reviews-bot

ldk-reviews-bot commented May 19, 2026

Copy link
Copy Markdown

👋 Thanks for assigning @jkczyz as a reviewer!
I'll wait for their review and will help manage the review process.
Once they submit their review, I'll check if a second reviewer would be helpful.

@ldk-claude-review-bot

ldk-claude-review-bot commented May 19, 2026

Copy link
Copy Markdown
Collaborator

I've completed a thorough re-review of the entire PR diff, examining all changes across channel.rs, channelmanager.rs, and splicing_tests.rs. I verified:

  1. Lock ordering safety: The new _with_funded_channel methods acquire short_to_chan_info.write(), pending_events.lock(), config.read(), and best_block.read() while peer_state is held. This follows the established lock ordering pattern already used throughout the codebase (e.g., in do_chain_event at lines 16379, 16395).

  2. splice_locked vs inferred_splice_locked separation: splice_locked (what we need to re-send) flows through FundingTxSignedhandle_channel_resumptionSendSpliceLocked message event. inferred_splice_locked (what we infer the peer sent) flows through internal_splice_locked_with_funded_channel for local promotion. These are correctly distinct paths.

  3. funding_locked_txid_sent_in_reestablish filtering: Correctly prevents duplicate splice_locked when the txid was already conveyed via my_current_funding_locked in the reestablish. The field is transient (not serialized), set at reestablish-send time, and consumed via .take() at reestablish-receive time.

  4. Ordering of splice promotion vs holding cell freeing: internal_splice_locked_with_funded_channel runs inside the match arm (before check_free_peer_holding_cells). If the resulting monitor update is async, the channel's pending monitor flag prevents maybe_free_holding_cell_htlcs from releasing cells prematurely.

  5. Error propagation: try_channel_entry! returns from the function on error, so check_free_peer_holding_cells is correctly skipped when the channel is force-closed.

  6. All constructor sites: All 5 ReestablishResponses sites include splice_locked, all 4 ChannelContext constructors include funding_locked_txid_sent_in_reestablish: None.

No issues found.

@codecov

codecovBot commented May 20, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 96.98795% with 5 lines in your changes missing coverage. Please review.
✅ Project coverage is 86.62%. Comparing base (1060865) to head (5e14a3f).
⚠️ Report is 12 commits behind head on main.

Files with missing linesPatch %Lines
lightning/src/ln/channelmanager.rs96.52%1 Missing and 4 partials ⚠️
Additional details and impacted files
@@ Coverage Diff @@## main #4624 +/- ##
==========================================
+ Coverage 86.59% 86.62% +0.03% 
==========================================
Files 159 159 Lines 110420 110568 +148 Branches 110420 110568 +148 ==========================================
+ Hits 95619 95784 +165 + Misses 12267 12250 -17 
Partials 2534 2534 
FlagCoverage Δ
fuzzing-fake-hashes6.61% <0.00%> (+0.04%)⬆️
fuzzing-real-hashes23.26% <28.91%> (+0.10%)⬆️
tests86.23% <96.98%> (+<0.01%)⬆️

Flags with carried forward coverage won't be shown. Click here to find out more.

☔ View full report in Codecov by Sentry.
📢 Have feedback on the report? Share it here.

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

Comment threadlightning/src/ln/channelmanager.rs
Comment threadlightning/src/ln/channelmanager.rs Outdated
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 16fb802 to 24bac54CompareMay 20, 2026 18:44
@wpaulino
wpaulino requested a review from jkczyzMay 20, 2026 18:44
Comment on lines +10524 to +10525
let funding_locked_txid_sent_in_reestablish =
self.context.funding_locked_txid_sent_in_reestablish.take();

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Is there any situation where we'd get here before calling get_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

No because get_channel_reestablish gets called as soon as the peer is tracked in the manager as connected

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13356 to +13375
if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready_with_peer_state(counterparty_node_id, &channel_ready_msg, peer_state)?;
}
};

self.handle_holding_cell_free_result(holding_cell_res);
// A reestablish may infer a missed `splice_locked`; apply it before freeing holding
// cells so we don't generate commitment updates against stale splice state.
let post_splice_locked_update = if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked_with_peer_state(counterparty_node_id, &splice_locked, peer_state)?
} else {
None
};

if let Some(channel_ready_msg) = need_lnd_workaround {
self.internal_channel_ready(counterparty_node_id, &channel_ready_msg)?;
}
let holding_cell_res = self.check_free_peer_holding_cells(peer_state);
(post_splice_locked_update, holding_cell_res)
};

if let Some(splice_locked) = inferred_splice_locked {
self.internal_splice_locked(counterparty_node_id, &splice_locked)?;
if let Some(data) = post_splice_locked_update {
self.handle_post_monitor_update_chan_resume(data);
}
self.handle_holding_cell_free_result(holding_cell_res);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Should we do most of this inside of the hash_map::Entry::Occupied arm above? Was thinking we'd avoid the duplicate lookups, too.

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Even just doing this change wasn't really worth it since this isn't a hot path

Comment threadlightning/src/ln/channelmanager.rs Outdated
Comment on lines +13607 to +13608
mem::drop(peer_state_lock);
mem::drop(per_peer_state);

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Why don't we need to do the same in internal_channel_reestablish?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Because there the peer state is declared within a nested scope

In most cases, we end up sending our `splice_locked` either implicitly
during reestablishment via
`ChannelReestablish::my_current_funding_locked`, or explicitly after
reestablishment. However, we did not consider that it's possible for the
node to be notified of the splice confirmation after connecting to their
peer but prior to reestablishing their channel. In such cases, we need
to explicitly send the `splice_locked` since it wasn't included in
`my_current_funding_locked`, but only after the channel has been
reestablished.
Found by the chanmon_consistency fuzz target.
Upon channel reestablishment, we free our holding cells to send any
pending updates to our peer. If we happened to implicitly lock a pending
splice during reestablishment, we want to make sure any updates we send
after the fact are considering the new channel state (post-splice), even
if the update was queued while the splice was still pending. Therefore,
we must always handle the inferred `splice_locked` first.
Found by the `chanmon_consistency` fuzz target.
@wpaulino
wpaulinoforce-pushed the splice-locked-reestablish-fixes branch from 24bac54 to 5e14a3fCompareMay 20, 2026 23:36
@jkczyz
jkczyz merged commit 9ce02b3 into lightningdevkit:mainMay 21, 2026
24 checks passed
@wpaulino
wpaulino deleted the splice-locked-reestablish-fixes branch May 22, 2026 20:03
@jkczyzjkczyz mentioned this pull request Jun 15, 2026
50 tasks
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

5 participants

@wpaulino@ldk-reviews-bot@ldk-claude-review-bot@TheBlueMatt@jkczyz