acl-org / acl-org/acl-anthology

Ingesting the NLP Highlights podcast

未关闭
#497 4 条评论 0 个 reaction 已指派 1 人 已被 @mbollmann 认领 在 GitHub 查看
ingestion
主要语言
Python
星标
796
派生
408
平均合并
3 天 13 小时
30 天内合并 PR
34

描述

The [NLP Highlights podcast](https://soundcloud.com/nlp-highlights) from AI2 has a wealth of material on a range of topics in NLP, from specific papers, to overviews of entire sub-fields, to career advice, and so on. I've listened to a number of these episodes and in addition to capturing a lot of interesting material, they are very well done. I got in touch with the hosts, Matt Gardner and Waleed Ammar, and AI2 is willing to archive this material in the Anthology, and also to contribute transcripts of the episodes.

I think this has lasting value and is worth preserving, but it presents two challenges that I'd like to solicit feedback on:

1. This is the first time we'll be ingesting non-peer-reviewed material
2. We currently only host papers as first-class objects

For (1), I don't think there is any issue, but it is probably worth discussing.

(2) is a technical issue related to how to represent this. NLP Highlights is currently not organized into anything like volumes. I have not given this too much thought. Perhaps we should have a second-level (beneath ``) `` tag, with third-level `` tags serving as paper analogs? We could use `` for the hosts and `` for whoever's being interviewed.

I think it would make sense to roll this out no sooner than EMLNP and no later than first quarter 2020, when we switch to the new Anthology ID format (#291).

Share your thoughts!

贡献指南

这个仓库没有索引到贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。