NLP Workbench: Efficient and Extensible Integration of State-of-the-art Text Mining Tools

要約

NLP Workbench は、専門家ではないユーザーが最先端のテキストマイニングモデルを使用して大規模なコーパスの意味を理解できるようにする、テキストマイニング用の Web ベースのプラットフォームです。
このプラットフォームは、学界の最新の事前トレーニング済みモデルとオープンソースシステムに基づいて構築されており、エンティティリンク、センチメント分析、セマンティック解析、関係抽出などのセマンティック分析機能を提供します。
その拡張可能な設計により、研究者や開発者は既存のモデルをスムーズに置き換えたり、新しいモデルを統合したりできます。
効率を向上させるために、アクセラレーションハードウェアの割り当てと計算の並列化を容易にするマイクロサービスアーキテクチャを採用しています。
このホワイトペーパーでは、NLP Workbench のアーキテクチャを紹介し、それを設計する際に直面した課題について説明します。
また、NLP Workbench のさまざまなユースケースと、それを他のアプローチよりも使用する利点についても説明します。
プラットフォームは活発に開発されており、そのソースコードは MIT ライセンスの下でリリースされています。
当社のプラットフォームを紹介するウェブサイトと短いビデオもご利用いただけます。

要約(オリジナル)

NLP Workbench is a web-based platform for text mining that allows non-expert users to obtain semantic understanding of large-scale corpora using state-of-the-art text mining models. The platform is built upon latest pre-trained models and open source systems from academia that provide semantic analysis functionalities, including but not limited to entity linking, sentiment analysis, semantic parsing, and relation extraction. Its extensible design enables researchers and developers to smoothly replace an existing model or integrate a new one. To improve efficiency, we employ a microservice architecture that facilitates allocation of acceleration hardware and parallelization of computation. This paper presents the architecture of NLP Workbench and discusses the challenges we faced in designing it. We also discuss diverse use cases of NLP Workbench and the benefits of using it over other approaches. The platform is under active development, with its source code released under the MIT license. A website and a short video demonstrating our platform are also available.

arxiv情報

著者	Peiran Yao,Matej Kosmajac,Abeer Waheed,Kostyantyn Guzhva,Natalie Hervieux,Denilson Barbosa
発行日	2023-03-02 16:59:31+00:00
arxivサイト	arxiv_id(pdf)

提供元, 利用サービス

arxiv.jp, Google

NLP Workbench: Efficient and Extensible Integration of State-of-the-art Text Mining Tools

要約

要約(オリジナル)

arxiv情報

提供元, 利用サービス

最近の投稿

最近のコメント

アーカイブ

カテゴリー