google / google/langextract

The langextract tool takes too long to locate the original content. Would it be possible to add a parameter to provide two options: one with original content localization and one without, so that it only extracts the content?

Open
#105 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
38.6k
Forks
2.7k
Avg merge
3d 15h
Merged PRs (30d)
3

Description

The langextract tool takes too long to locate the original content. Would it be possible to add a parameter to provide two options: one with original content localization and one without, so that it only extracts the content?

Contributor guide

Open the contributing guide

Research direction

No files or tests are named in the issue. Start by locating the langextract tool entry point and the code that performs original-content localization; determine how an option can bypass that step while retaining extraction, then add coverage showing both modes behave as intended.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Feature
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.