All skills
brianlovin avatar

/fix-sentry-issues

@371c6c5 official

Use Sentry MCP to discover, triage, and fix production issues with root-cause analysis. Use when asked to fix Sentry issues, triage production errors, investigate error spikes, or clean up Sentry noise. Requires Sentry MCP server. Triggers on "fix sentry", "triage errors", "production bugs", "sentry issues".

Use this Skill: https://skilld.dev/gh/brianlovin/claude-config/fix-sentry-issues

This session only. Nothing lands on disk.

SKILL.md

≈82 tokens always: the name and description. ≈1.1k when used: this file.

Fix Sentry Issues

Philosophy

The Sentry error is not the problem. It's a signal.

Your goal is not to close the Sentry issue. Your goal is to discover the root cause, understand what's wrong with the application, and fix the underlying defect. Closing the Sentry issue is a side effect of doing that correctly.

Ask "Why does this fail?" — not "How do I make Sentry quiet?" Never treat log level changes as fixes. A fallback path means degraded user experience; trace why the primary path fails and fix it upstream.

Anti-patterns (do not do these)

  • Batch-classifying as "expected" without investigation. Seeing a fallback does NOT mean you understand the failure. Trace the full input path.
  • Treating "has a fallback" as "not a problem." Why does the primary path fail? Can we prevent it upstream?
  • Combining multiple issues into one PR. Each has its own root cause. Fix individually (except when investigation proves identical cause).
  • Throwing away error details. Never remove error from catch (error) or strip status codes. That data is how you understand failures.
  • Deciding the fix during triage. Classify as "Investigate" or "Ignore" only. You don't know the fix until investigation is complete.

Log level downgrade is valid ONLY for genuinely expected states (e.g., optional column missing, resource deleted) — NOT for failures with fallbacks.

Phase 1: Discover & Triage

Use Sentry MCP (ToolSearch first to load tools): find_organizations → find_projects → search_issues with naturalLanguageQuery: "all unresolved issues sorted by events".

Build a triage table. Action = Investigate or Ignore only:

ID Title Events Action Reason
PROJ-A Error in save 14 Investigate User-facing save failure
PROJ-B GM_register... 3 Ignore Greasemonkey extension

Investigate: multiple events, degraded user experience, high-volume warnings, recurring on every run. Ignore: browser extension code, ChunkLoadError (self-resolving), single-event transients, already fixed.

Apply: mcp__sentry__update_issue(..., status: "ignored") or status: "resolved" for already-fixed.

Phase 2: Investigate (one issue at a time)

Work through these steps in order. Do not skip or batch issues.

  1. Pull event-level data — Issue summaries hide details. Use get_issue_details and search_issue_events with naturalLanguageQuery: "all events with extra data". Extract: URLs, params, stack traces, status codes, timestamps.

  2. Cross-reference Axiom — Events have traceId. axiom query "['shiori-events'] | where traceId == '<traceId>'" -f json for surrounding context (authMethod, client_version, request metadata).

  3. Read the failing code path — Follow the stack trace. Read every file. Understand before proposing changes.

  4. Trace the input path upstream (most often skipped, most important) — What data reaches the failing function? Should it have reached this path at all? Is there a missing filter? Is the input wrong (binary URL, redirect, bad format)? Can we prevent bad inputs upstream?

  5. Reproduce — Use actual failing inputs from Sentry. Call the function with exact data. fetch() the URLs that timed out. Verify your understanding.

  6. Identify root cause — Why does this input fail? Why does it reach this path? What's the right fix? (e.g., "Filter binary URLs before Firecrawl" — not "suppress the log")

Pattern Real Fix
External API fails on certain URLs Filter/validate inputs before sending
Timeout Investigate what's slow; adjust timeout or input size
DB "invalid json" Sanitize before insert
Stale reference on cron Detect staleness, auto-clean

Phase 3: Fix

One branch per issue. git checkout main && git pull && git checkout -b fix/<descriptive-name>

  • Tests first — Use data from actual Sentry events. Test fails before fix, passes after.
  • Implement — Fix the root cause, not the symptom. If the fix is primarily a log level change, STOP: did you investigate why it fails, or just suppress?
  • Verify — Tests pass, lint passes, fix handles actual failing inputs.
  • PR — Include Root cause (upstream reason) and Fix (what changed and why it prevents the failure). Resolve in Sentry only after merge.

Source: SKILL.md on GitHub

1 warning16d5 checks · Risk SAFE
  • Gen Agent Trust Hub16d

    The skill is designed to help triage and fix production errors using the Sentry MCP server and Axiom queries. It is generally safe but exposes an indirect prompt injection surface because it processes untrusted data from Sentry logs, error details, and stack traces, and instructs the agent to execute local commands and perform HTTP requests to URLs found in logs.

  • Socket16d

    No alerts

  • Snyk16d

    Risk: MEDIUM · 1 issue

  • Runlayer7mo

    1/1 file flagged

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at 371c6c5. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub 2 months ago.

Dormantupdated 7 months ago
  • Debugging
  • MCP
  • sentry
  • error-tracking
  • production
  • root-cause-analysis
  • triage
  • monitoring

README badge

README badge for brianlovin/claude-config/fix-sentry-issues

Connects to a Sentry MCP server to discover, triage, and fix production errors by tracing root causes rather than suppressing symptoms. Use when investigating error spikes, unresolved issues, or production failures—the skill enforces investigation of input paths and upstream sources before proposing fixes.

Generated from the current SKILL.md.

What does this skill require to run?
A Sentry MCP server must be configured and running. The skill uses the Sentry MCP's tools (find_organizations, search_issues, get_issue_details, update_issue) to query and triage production errors.
Does this skill automatically fix issues?
No. The skill guides investigation and root-cause analysis across three phases (discover, investigate, fix), but the actual code changes and PR creation are your responsibility after identifying the underlying defect.
Can I use this to suppress errors by downgrading log levels?
Only if the error represents a genuinely expected state (optional field missing, resource already deleted). The skill explicitly warns against using log level changes to hide failures with fallbacks; you must trace and fix the upstream cause instead.
How does this handle multiple Sentry issues at once?
Each issue gets triaged individually as either Investigate or Ignore. During investigation and fixes, the skill processes one issue at a time to avoid conflating separate root causes into a single PR.
Does this integrate with Axiom or other observability tools?
The skill references Axiom queries using traceId from Sentry events for surrounding context (auth method, client version, request metadata), but the integration depends on your Axiom setup and MCP configuration.

Generated from the current SKILL.md. These answers refresh after source changes.