perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Add copy buttons to all
 blocks\n(function() {\n function addCopyButtons() {\n document.querySelectorAll('pre code').forEach(function(codeBlock) {\n if (codeBlock.parentElement.hasAttribute('data-copy-added')) return;\n codeBlock.parentElement.setAttribute('data-copy-added', 'true');\n \n var btn = document.createElement('button');\n btn.textContent = 'Copy';\n btn.style.cssText = 'position:absolute;top:4px;right:4px;padding:2px 8px;font-size:11px;background:#4ecdc4;border:none;border-radius:4px;color:#1a1a2e;cursor:pointer;opacity:0.7;transition:opacity 0.2s;';\n btn.onmouseover = function() { this.style.opacity = '1'; };\n btn.onmouseout = function() { this.style.opacity = '0.7'; };\n btn.onclick = function() {\n navigator.clipboard.writeText(codeBlock.textContent).then(function() {\n btn.textContent = 'Copied!';\n setTimeout(function() { btn.textContent = 'Copy'; }, 1500);\n });\n };\n codeBlock.parentElement.style.position = 'relative';\n codeBlock.parentElement.appendChild(btn);\n });\n }\n \n addCopyButtons();\n \n // Re-run on dynamic content\n var observer = new MutationObserver(addCopyButtons);\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Add Copy Buttons to Code Blocks");
}
} catch(__e) { console.warn('[Userscript:Add Copy Buttons to Code Blocks]', __e); }
})();
(function(){
try {
var __m = "github.com";
var __re = new RegExp('^' + "github\\.com" + '
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Force GitHub README to respect dark mode\n(function() {\n var style = document.createElement('style');\n style.textContent = '\n .markdown-body {\n color-scheme: dark light;\n }\n .markdown-body pre { background: #161b22 !important; }\n .markdown-body code { background: rgba(110, 118, 129, 0.4) !important; }\n .markdown-body table th, .markdown-body table td { border-color: #30363d !important; }\n .markdown-body img { background: #0d1117; }\n .markdown-body blockquote { border-left-color: #8b949e; }\n .markdown-body hr { border-color: #30363d; }\n ';\n document.head.appendChild(style);\n})();", "GitHub Dark Mode README Fix"); } } catch(__e) { console.warn('[Userscript:GitHub Dark Mode README Fix]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Highlight search terms from Google/DuckDuckGo/Bing referrer\n(function() {\n var ref = document.referrer;\n var terms = [];\n \n if (ref.includes('google.com') || ref.includes('duckduckgo.com') || ref.includes('bing.com')) {\n var url = new URL(ref);\n var q = url.searchParams.get('q') || url.searchParams.get('p');\n if (q) {\n terms = q.split(/\\s+/).filter(function(t) { return t.length > 2; });\n }\n }\n \n if (terms.length === 0) return;\n \n var style = document.createElement('style');\n style.textContent = '.userscript-highlight { background: #fbbf24; color: #1a1a2e; padding: 1px 3px; border-radius: 2px; }';\n document.head.appendChild(style);\n \n function highlight(node) {\n if (node.nodeType === 3) { // text node\n var text = node.textContent;\n var found = false;\n terms.forEach(function(term) {\n var regex = new RegExp('(' + term.replace(/[.*+?^${}()|[\\]\\\\]/g, '\\\\') + ')', 'gi');\n if (regex.test(text)) {\n found = true;\n var frag = document.createDocumentFragment();\n var parts = text.split(regex);\n parts.forEach(function(part, i) {\n if (i % 2 === 0) {\n frag.appendChild(document.createTextNode(part));\n } else {\n var span = document.createElement('span');\n span.className = 'userscript-highlight';\n span.textContent = part;\n frag.appendChild(span);\n }\n });\n node.parentNode.replaceChild(frag, node);\n }\n });\n } else if (node.nodeType === 1 && node.childNodes) { // element\n var skipTags = ['SCRIPT', 'STYLE', 'NOSCRIPT', 'TEXTAREA', 'INPUT', 'SELECT'];\n if (!skipTags.includes(node.tagName)) {\n Array.from(node.childNodes).forEach(highlight);\n }\n }\n }\n \n highlight(document.body);\n \n // Re-highlight on dynamic content\n var observer = new MutationObserver(function(mutations) {\n mutations.forEach(function(m) {\n m.addedNodes.forEach(function(node) {\n if (node.nodeType === 1 || node.nodeType === 3) highlight(node);\n });\n });\n });\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Highlight Search Terms"); } } catch(__e) { console.warn('[Userscript:Highlight Search Terms]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Strip utm_, fbclid, gclid, etc. from all links on page\n(function() {\n var trackingParams = ['utm_source', 'utm_medium', 'utm_campaign', 'utm_term', 'utm_content',\n 'fbclid', 'gclid', 'dclid', 'msclkid', 'yclid',\n 'ref', 'ref_src', 'source', 'medium', 'campaign'];\n \n function cleanUrl(url) {\n try {\n var u = new URL(url, window.location.origin);\n var changed = false;\n trackingParams.forEach(function(p) {\n if (u.searchParams.has(p)) {\n u.searchParams.delete(p);\n changed = true;\n }\n });\n return changed ? u.toString() : url;\n } catch (e) {\n return url;\n }\n }\n \n function cleanLinks() {\n document.querySelectorAll('a[href]').forEach(function(a) {\n var clean = cleanUrl(a.href);\n if (clean !== a.href) a.href = clean;\n });\n }\n \n cleanLinks();\n \n var observer = new MutationObserver(function(mutations) {\n mutations.forEach(function(m) {\n m.addedNodes.forEach(function(node) {\n if (node.nodeType === 1) {\n if (node.tagName === 'A') cleanLinks();\n node.querySelectorAll('a[href]').forEach(function(a) {\n var clean = cleanUrl(a.href);\n if (clean !== a.href) a.href = clean;\n });\n }\n });\n });\n });\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Remove Tracking Parameters from Links"); } } catch(__e) { console.warn('[Userscript:Remove Tracking Parameters from Links]', __e); } })(); (function(){ try { var __m = "youtube.com"; var __re = new RegExp('^' + "youtube\\.com" + '
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Auto-enable theater mode on YouTube\n(function() {\n function tryTheater() {\n var btn = document.querySelector('button[aria-label=\"Theater mode\"], ytd-player #player button[title=\"Theater mode\"]');\n if (btn && !btn.classList.contains('activated')) {\n btn.click();\n }\n }\n \n // Try immediately\n tryTheater();\n \n // Try after navigation (SPA)\n var lastUrl = location.href;\n setInterval(function() {\n if (location.href !== lastUrl) {\n lastUrl = location.href;\n setTimeout(tryTheater, 500);\n }\n }, 1000);\n \n // Also try on player load\n var observer = new MutationObserver(tryTheater);\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "YouTube Theater Mode Default"); } } catch(__e) { console.warn('[Userscript:YouTube Theater Mode Default]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Remove or un-stick sticky/fixed headers that block content\n(function() {\n function unstick() {\n document.querySelectorAll('header, nav, [role=\"banner\"], .header, .navbar, .sticky, .fixed-top, [style*=\"position: fixed\"], [style*=\"position:sticky\"]').forEach(function(el) {\n if (el.style.position === 'fixed' || el.style.position === 'sticky' || \n getComputedStyle(el).position === 'fixed' || getComputedStyle(el).position === 'sticky') {\n el.style.position = 'static';\n el.style.top = 'auto';\n el.style.zIndex = 'auto';\n }\n });\n }\n \n unstick();\n \n var observer = new MutationObserver(unstick);\n observer.observe(document.body, { childList: true, subtree: true, attributes: true, attributeFilter: ['style', 'class'] });\n})();", "Kill Sticky Headers"); } } catch(__e) { console.warn('[Userscript:Kill Sticky Headers]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Universal Dark Mode - works on any site\n(function() {\n var enabled = true;\n \n function applyDarkMode() {\n if (!enabled) return;\n \n // Create style element if it doesn't exist\n var style = document.getElementById('universal-dark-mode-style');\n if (!style) {\n style = document.createElement('style');\n style.id = 'universal-dark-mode-style';\n document.head.appendChild(style);\n }\n \n // Dark mode CSS - inverts colors but preserves images/video\n style.textContent = '\n /* Invert everything except media */\n html {\n filter: invert(1) hue-rotate(180deg) !important;\n background: #1a1a2e !important;\n }\n \n /* Restore images, videos, iframes, canvas */\n img, video, iframe, canvas, svg, picture, [style*=\"background-image\"] {\n filter: invert(1) hue-rotate(180deg) !important;\n }\n \n /* Preserve specific elements that should not be inverted */\n .no-dark-mode, .no-dark-mode *,\n [data-theme=\"light\"], [data-theme=\"light\"],\n .ace_editor, .ace_editor *,\n .CodeMirror, .CodeMirror *,\n .monaco-editor, .monaco-editor *,\n .markdown-body pre, .markdown-body pre *,\n .highlight, .highlight *,\n pre code, pre code * {\n filter: none !important;\n }\n \n /* Fix common UI elements */\n .modal, .popup, .dropdown-menu, .tooltip, .popover {\n filter: invert(1) hue-rotate(180deg) !important;\n background: #2d2d44 !important;\n border-color: #444 !important;\n }\n \n /* Scrollbars */\n ::-webkit-scrollbar { background: #1a1a2e !important; }\n ::-webkit-scrollbar-thumb { background: #444 !important; }\n ::-webkit-scrollbar-thumb:hover { background: #555 !important; }\n \n /* Selection */\n ::selection { background: #4ecdc4 !important; color: #1a1a2e !important; }\n ::-moz-selection { background: #4ecdc4 !important; color: #1a1a2e !important; }\n ';\n }\n \n function removeDarkMode() {\n var style = document.getElementById('universal-dark-mode-style');\n if (style) style.remove();\n }\n \n // Toggle with Alt+Shift+D\n document.addEventListener('keydown', function(e) {\n if (e.altKey && e.shiftKey && e.key === 'D') {\n e.preventDefault();\n enabled = !enabled;\n if (enabled) {\n applyDarkMode();\n console.log('[Universal Dark Mode] Enabled');\n } else {\n removeDarkMode();\n console.log('[Universal Dark Mode] Disabled');\n }\n }\n });\n \n // Apply on load\n applyDarkMode();\n \n // Re-apply on dynamic content\n var observer = new MutationObserver(function(mutations) {\n if (enabled && !document.getElementById('universal-dark-mode-style')) {\n applyDarkMode();\n }\n });\n observer.observe(document.head, { childList: true });\n \n console.log('[Universal Dark Mode] Loaded - Press Alt+Shift+D to toggle');\n})();", "Universal Dark Mode"); } } catch(__e) { console.warn('[Userscript:Universal Dark Mode]', __e); } })(); })();
Skip to content

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178) - #244

Closed
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes
Closed

perf(db): add missing hot-path indexes (#160, #161, #176, #177, #178)#244
Jose-Gael-Cruz-Lopez wants to merge 24 commits into
mainfrom
perf/db-indexes

Conversation

@Jose-Gael-Cruz-Lopez

Copy link
Copy Markdown
Member

Adds the ten missing indexes called out in the backend performance audit. Every query site listed below currently full-scans + sorts the target table on the critical path.

Indexes added

IssueTableIndex
#161messages(session_id, created_at) — chat-history load (fastest-growing table)
#160graph_edges(user_id), (source_node_id), (target_node_id) — graph render + cascade delete
#176sessions(user_id, started_at DESC) — history list + profile stats
#177documents(user_id, created_at DESC), (user_id, course_id) — library list + study-guide context
#178study_guides(user_id, generated_at DESC), (user_id, course_id, exam_id) — guide list + cache lookup
#178quiz_attempts(user_id) — achievement counts + history aggregations

Changes

  • backend/db/migration_perf_indexes.sql — new hand-applied migration (all CREATE INDEX IF NOT EXISTS, re-runnable).
  • backend/db/supabase_schema.sql — same indexes mirrored so fresh environments match prod.
  • backend/tests/test_perf_indexes_present.py — drift guard asserting each index lives in both files and is IF NOT EXISTS-guarded.

Verification

  • ruff check . clean; full gated suite green (685 + 3 new).
  • No local Postgres in CI, so EXPLAIN confirmation is left to a reviewer with DB access; index column orders match the documented query filters/sorts.

Closes#160, #161, #176, #177, #178.

@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 22, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitUpdated (UTC)
❌ Deployment failed
View logs
frontend28ae402Jun 22 2026, 03:52 AM

@coderabbitai

coderabbitaiBot commented Jun 22, 2026

Copy link
Copy Markdown

Warning

Review limit reached

@Jose-Gael-Cruz-Lopez, we couldn't start this review because you've reached your PR review rate limit.

More reviews will be available in 47 minutes and 35 seconds. Learn how PR review limits work.

Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file).

⌛ How to resolve this issue?

After more reviews become available, a review can be triggered using the @coderabbitai review command as a PR comment. Alternatively, push new commits to this PR.

To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits.

🚦 How do rate limits work?

CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability.

For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window.

Please see our Fair Usage Limits Policy for further information.

ℹ️ Review info
⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: cc582818-ee47-4199-bfe2-5ada89bb2a3d

📥 Commits

Reviewing files that changed from the base of the PR and between 32409a8 and 4d1e6a1.

📒 Files selected for processing (3)
  • backend/db/migrations/0001_baseline_schema.sql
  • backend/db/migrations/0019_perf_indexes.sql
  • backend/tests/test_perf_indexes_present.py
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch perf/db-indexes

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

The hand-applied migration ran plain CREATE INDEX, which takes an
ACCESS EXCLUSIVE lock on each table for the whole build and would stall
live traffic on hot tables (messages, sessions, graph_edges, documents).
Switch every statement to CREATE INDEX CONCURRENTLY IF NOT EXISTS and
document the non-transactional / INVALID-index recovery caveats. The
canonical schema keeps plain CREATE INDEX since it runs on empty tables.
idx_graph_edges_source/target are single-column rather than composite
with user_id on purpose: the bulk dedup deletes in db/dedup_nodes.py
filter the endpoint via in.(...) without user_id, which a leading-user_id
composite could not serve. Note the rationale in the schema.
The drift guard only checked that each table name appeared somewhere in
the migration, which a comment mention would satisfy. Match the index
name followed by an ON <table>( clause instead so a misdirected index is
caught.
@cloudflare-workers-and-pages

cloudflare-workers-and-pagesBot commented Jun 24, 2026

Copy link
Copy Markdown

Deploying with Cloudflare Workers Cloudflare Workers

The latest updates on your project. Learn more about integrating Git with Workers.

StatusNameLatest CommitPreview URLUpdated (UTC)
✅ Deployment successful!
View logs
frontend-staging4d1e6a1Commit Preview URL

Branch Preview URL
Jun 24 2026, 02:48 PM

The flat db/migration_perf_indexes.sql predates main's ordered
migrations/ layout (run by db/migrate.py) and was never picked up by the
runner. Replace it with a numbered, idempotent migration so existing
databases (baselined before these indexes existed) get them. Fresh DBs
already get them via 0001_baseline_schema.sql.
Uses plain CREATE INDEX IF NOT EXISTS instead of CONCURRENTLY because
migrate.py applies each migration inside a transaction block.
The test read supabase_schema.sql and migration_perf_indexes.sql, both
gone after main restructured db files. Read the canonical locations that
exist now: 0001_baseline_schema.sql (fresh DBs) and 0019_perf_indexes.sql
(existing DBs), keeping the on-table + idempotency assertions.
@AndresL230

Copy link
Copy Markdown
Collaborator

Superseded by the DB modular redesign (#279). All of these indexes (#160/#161/#176/#177/#178) ship in migrations 0023/0025. Closing as obsolete — this migration would collide with the redesign. Reopen if anything here isn't covered by #279.

@AndresL230
AndresL230 deleted the perf/db-indexes branch June 27, 2026 04:21
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

graph_edges has zero indexes — every knowledge-graph render and node delete full-scans

2 participants

@Jose-Gael-Cruz-Lopez@AndresL230