Skip to content
Back to skills

Bug Investigation 1

ASecurity

Systematic bug investigation and root cause analysis

  • 2 stars
  • 0 votes
  • 0 copies
  • 0 views
  • Added September 27, 2026
testingpythondebuggingdatabaseperformance

Security analysis

A100/100

Scanned September 27, 2026

npx -y skills add David-Li0406/meta-skill-evloving --skill bug-investigation-1 --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Bug Investigation 1?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Bug Investigation 1
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/david-li0406-bug-investigation-1/badge)](https://www.skillsdirectory.com/skills/david-li0406-bug-investigation-1)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
type: skill
name: Bug Investigation
description: Systematic bug investigation and root cause analysis
skillSlug: bug-investigation
phases: [E, V]
generated: 2026-01-20
status: filled
scaffoldVersion: "2.0.0"
---

# Bug Investigation Skill

## When to Use

Use this skill when:
- Debugging extraction failures
- Investigating classification errors
- Analyzing search performance issues
- Troubleshooting database problems

## Debugging Workflow

### 1. Reproduce the Issue

- Identify failing PDF or operation
- Reproduce in isolation
- Collect error messages

### 2. Check Extraction Metrics

Use `metrics.py` to check:
- Extraction success rate
- Methods used (pdfplumber vs OCR)
- Error patterns

### 3. Review Logs

- Error messages in console
- Database error logs
- Processing statistics

## Common Bug Patterns

### PDF Extraction Failures

**Symptoms:**
- "Sem texto extraível"
- Empty content in database
- OCR not triggered when needed

**Investigation:**
1. Check if PDF is scanned (images)
2. Verify OCR is installed and working
3. Test extraction manually
4. Check file permissions

### Classification Errors

**Symptoms:**
- Documents classified as "outros"
- Incorrect contract number extraction
- Missing document numbers

**Investigation:**
1. Check filename pattern
2. Test regex patterns
3. Verify classification logic
4. Review expected vs actual output

### Database Issues

**Symptoms:**
- Duplicate key errors
- FTS5 index not updating
- Missing data in results

**Investigation:**
1. Check filepath uniqueness
2. Verify triggers are working
3. Test queries directly
4. Check database schema

### Search Problems

**Symptoms:**
- No results found
- Incorrect results
- Performance issues

**Investigation:**
1. Verify FTS5 index exists
2. Test query syntax
3. Check content was indexed
4. Review filter logic

## Logging and Error Handling

### Error Logging Pattern

```python
try:
    text = extract_text_from_pdf(full_path)
except Exception as e:
    print(f"   ❌ Erro ao processar {file}: {e}")
    # Log error with context
    errors += 1
```

### Debugging Checklist

- [ ] Error message clear and helpful
- [ ] Context information logged
- [ ] Error doesn't crash entire process
- [ ] Metrics track failures

## Test Verification Steps

### For Extraction Bugs

1. Test with sample PDF
2. Verify extraction method used
3. Check text length
4. Validate OCR if used

### For Classification Bugs

1. Test classification function directly
2. Verify regex matches
3. Check fallback logic
4. Compare with expected result

## Bug Fix Examples

### Fix: OCR Not Triggering

**Root Cause:** OCR check happens after text validation

**Fix:** Move OCR check before validation failure

### Fix: Classification Fails

**Root Cause:** Regex doesn't match all patterns

**Fix:** Improve regex or add alternative patterns

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…