Show simple item record

dc.contributor.authorLison, Pierre
dc.contributor.authorBarnes, Jeremy
dc.contributor.authorHubin, Aliaksandr
dc.date.accessioned2021-08-12T06:14:26Z
dc.date.available2021-08-12T06:14:26Z
dc.date.created2021-08-11T15:38:23Z
dc.date.issued2021
dc.identifier.isbn978-1-954085-56-5
dc.identifier.urihttps://hdl.handle.net/11250/2767430
dc.description.abstractWe present skweak, a versatile, Python-based software toolkit enabling NLP developers to apply weak supervision to a wide range of NLP tasks. Weak supervision is an emerging machine learning paradigm based on a simple idea: instead of labelling data points by hand, we use labelling functions derived from domain knowledge to automatically obtain annotations for a given dataset. The resulting labels are then aggregated with a generative model that estimates the accuracy (and possible confusions) of each labelling function. The skweak toolkit makes it easy to implement a large spectrum of labelling functions (such as heuristics, gazetteers, neural models or linguistic constraints) on text data, apply them on a corpus, and aggregate their results in a fully unsupervised fashion. skweak is especially designed to facilitate the use of weak supervision for NLP tasks such as text classification and sequence labelling. We illustrate the use of skweak for NER and sentiment analysis. skweak is released under an open-source license and is available at https://github.com/NorskRegnesentral/skweak
dc.language.isoengen_US
dc.relation.ispartofProceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing: System Demonstrations
dc.relation.urihttps://github.com/NorskRegnesentral/skweak
dc.rightsNavngivelse-Ikkekommersiell-DelPåSammeVilkår 4.0 Internasjonal*
dc.rights.urihttp://creativecommons.org/licenses/by-nc-sa/4.0/deed.no*
dc.titleskweak: Weak Supervision Made Easy for NLPen_US
dc.typeChapteren_US
dc.description.versionpublishedVersion
cristin.ispublishedtrue
cristin.fulltextoriginal
cristin.qualitycode1
dc.identifier.cristin1925385
dc.source.pagenumber337-346en_US
dc.relation.projectNotur/NorStore: NN9850K
dc.relation.projectNorges forskningsråd: 300921
dc.relation.projectNorges forskningsråd: 308904


Files in this item

Thumbnail

This item appears in the following Collection(s)

Show simple item record

Navngivelse-Ikkekommersiell-DelPåSammeVilkår 4.0 Internasjonal
Except where otherwise noted, this item's license is described as Navngivelse-Ikkekommersiell-DelPåSammeVilkår 4.0 Internasjonal