acl-org / acl-org/reviewer-paper-matching

do you want to use archived Semantic Scholar data from the API?

Open
#17 0 comments 0 reactions 1 assignee Assigned to @ajstent View on GitHub
Dominant language
Python
Stars
26
Forks
4
PR merge metrics
No merged PRs in 30d

Description

the checkProfiles and collectCOIs scripts in the coi repo now archive all the queries they make to Semantic Scholar (in a authors.json and papers.json file respectively). So you could use this data directly to similarly speed up reviewer assignment (and the data is much more up to date than the latest dump). I attach sample files (both are in the attached zip) from the larger profile sample csv we have been using for testing.

[authors.zip](https://github.com/acl-org/reviewer-paper-matching/files/3906258/authors.zip)

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reading the checkProfiles and collectCOIs scripts and inspecting the attached authors.zip, including authors.json and papers.json. Determine whether reviewer assignment can consume these archives instead of issuing the existing queries; done means an agreed implementation scope and a demonstrable speedup using the supplied sample data.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data, performance
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.