Create benchmark of real-world attacks and large codebases
Open
- Dominant language
- Python
- Stars
- 63
- Forks
- 13
- PR merge metrics
- No merged PRs in 30d
Description
On `dev` several detectors give too many false positives. We need to create a benchmark to get a better sense of how to improve the current detectors and make sure tealer give a good rate of FP <> FN <> TP
Related: https://github.com/crytic/tealer/issues/65
Contributor guide
Research direction
Start by reviewing the detectors on the `dev` branch and the discussion in related issue #65. Define a benchmark containing real-world attacks and large codebases, then measure false positives, false negatives, and true positives to determine when the work is complete.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- security, testing-qa
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100