| Segment Any Text: A Universal Approach for Robust, Efficient and Adaptable Sentence Segmentation | Jun 24, 2024 | parameter-efficient fine-tuningSentence | CodeCode Available | 7 | 5 |
| Where's the Point? Self-Supervised Multilingual Punctuation-Agnostic Sentence Segmentation | May 30, 2023 | Machine TranslationSegmentation | CodeCode Available | 3 | 5 |
| Abstractive Summarization of Spoken andWritten Instructions with BERT | Aug 21, 2020 | Abstractive Text SummarizationArticles | CodeCode Available | 2 | 5 |
| Trankit: A Light-Weight Transformer-based Toolkit for Multilingual Natural Language Processing | Jan 9, 2021 | Dependency ParsingLanguage Modeling | CodeCode Available | 1 | 5 |
| Opera Graeca Adnotata: Building a 34M+ Token Multilayer Corpus for Ancient Greek | Mar 31, 2024 | LemmatizationSentence | CodeCode Available | 1 | 5 |
| A unified approach to sentence segmentation of punctuated text in many languages | Aug 1, 2021 | SentenceSentence segmentation | CodeCode Available | 1 | 5 |
| Not Low-Resource Anymore: Aligner Ensembling, Batch Filtering, and New Datasets for Bengali-English Machine Translation | Sep 20, 2020 | Machine TranslationSentence | CodeCode Available | 1 | 5 |
| Mukayese: Turkish NLP Strikes Back | Mar 2, 2022 | BenchmarkingLanguage Modeling | CodeCode Available | 1 | 5 |
| Ascle: A Python Natural Language Processing Toolkit for Medical Text Generation | Nov 28, 2023 | Machine TranslationQuestion Answering | CodeCode Available | 1 | 5 |
| Lexical Semantic Recognition | Apr 30, 2020 | Natural Language UnderstandingSentence | CodeCode Available | 1 | 5 |
| KG-GPT: A General Framework for Reasoning on Knowledge Graphs Using Large Language Models | Oct 17, 2023 | Fact VerificationKnowledge Graphs | CodeCode Available | 1 | 5 |
| Abstractive Summarization of Spoken and Written Instructions with BERT | Aug 21, 2020 | Abstractive Text SummarizationArticles | CodeCode Available | 1 | 5 |
| Towards JointUD: Part-of-speech Tagging and Lemmatization using Recurrent Neural Networks | Sep 10, 2018 | Dependency ParsingLemmatization | CodeCode Available | 0 | 5 |
| Creating a Universal Dependencies Treebank of Spoken Frisian-Dutch Code-switched Data | Feb 22, 2021 | SentenceSentence segmentation | CodeCode Available | 0 | 5 |
| Evaluating Sentence Segmentation and Word Tokenization Systems on Estonian Web Texts | Nov 16, 2020 | SegmentationSentence | CodeCode Available | 0 | 5 |
| Human Genome Book: Words, Sentences and Paragraphs | Jan 23, 2025 | Protein Structure PredictionSentence segmentation | CodeCode Available | 0 | 5 |
| LeConTra: A Learner Corpus of English-to-Dutch News Translation | Jun 1, 2022 | SentenceSentence segmentation | CodeCode Available | 0 | 5 |
| Prosodic features improve sentence segmentation and parsing | Feb 23, 2023 | SentenceSentence segmentation | CodeCode Available | 0 | 5 |
| Fine-Grained Argument Unit Recognition and Classification | Apr 22, 2019 | Argument MiningArgument Retrieval | CodeCode Available | 0 | 5 |
| SLATE: A Sequence Labeling Approach for Task Extraction from Free-form Inked Content | Nov 8, 2022 | FormSegmentation | CodeCode Available | 0 | 5 |
| Universal Dependency Parsing from Scratch | Jan 29, 2019 | AllDependency Parsing | CodeCode Available | 0 | 5 |
| Using Punkt for Sentence Segmentation in non-Latin Scripts: Experiments on Kurdish (Sorani) Texts | Apr 9, 2020 | SentenceSentence segmentation | CodeCode Available | 0 | 5 |
| Fine-Grained Control of Sentence Segmentation and Entity Positioning in Neural NLG | Nov 1, 2019 | Data-to-Text GenerationPosition | —Unverified | 0 | 0 |
| From Raw Text to Universal Dependencies - Look, No Tags! | Aug 1, 2017 | Dependency ParsingPart-Of-Speech Tagging | —Unverified | 0 | 0 |
| GujiBERT and GujiGPT: Construction of Intelligent Information Processing Foundation Language Models for Ancient Texts | Jul 11, 2023 | Model SelectionPart-Of-Speech Tagging | —Unverified | 0 | 0 |
| Transformer-Encoder-GRU (T-E-GRU) for Chinese Sentiment Analysis on Chinese Comment Text | Aug 1, 2021 | Chinese Sentiment AnalysisPosition | —Unverified | 0 | 0 |
| IBM Research at the CoNLL 2018 Shared Task on Multilingual Parsing | Oct 1, 2018 | ARCDependency Parsing | —Unverified | 0 | 0 |
| IMS at the CoNLL 2017 UD Shared Task: CRFs and Perceptrons Meet Neural Networks | Aug 1, 2017 | POSSegmentation | —Unverified | 0 | 0 |
| Inforex -- a web-based tool for text corpus management and semantic annotation | May 1, 2012 | ManagementNamed Entity Recognition (NER) | —Unverified | 0 | 0 |
| Integration of Automatic Sentence Segmentation and Lexical Analysis of Ancient Chinese based on BiLSTM-CRF Model | May 1, 2020 | Lexical Analysisnamed-entity-recognition | —Unverified | 0 | 0 |
| Creating Training Corpora for NLG Micro-Planners | Jul 1, 2017 | Data-to-Text GenerationReferring Expression | —Unverified | 0 | 0 |
| UDPipe 2.0 Prototype at CoNLL 2018 UD Shared Task | Oct 1, 2018 | Dependency ParsingLemmatization | —Unverified | 0 | 0 |
| Corpus Augmentation by Sentence Segmentation for Low-Resource Neural Machine Translation | May 22, 2019 | Low Resource Neural Machine TranslationLow-Resource Neural Machine Translation | —Unverified | 0 | 0 |
| Classical Chinese Sentence Segmentation for Tomb Biographies of Tang Dynasty | Aug 28, 2019 | BIG-bench Machine LearningSentence | —Unverified | 0 | 0 |
| Midas Loop: A Prioritized Human-in-the-Loop Annotation for Large Scale Multilayer Data | Jun 1, 2022 | Active LearningManagement | —Unverified | 0 | 0 |
| Mukayese: Turkish NLP Strikes Back | Nov 16, 2021 | BenchmarkingLanguage Modeling | —Unverified | 0 | 0 |
| Bulgarian X-language Parallel Corpus | May 1, 2012 | named-entity-recognitionNamed Entity Recognition | —Unverified | 0 | 0 |
| Better Chinese Sentence Segmentation with Reinforcement Learning | Aug 1, 2021 | reinforcement-learningReinforcement Learning | —Unverified | 0 | 0 |
| Online Sentence Segmentation for Simultaneous Interpretation using Multi-Shifted Recurrent Neural Network | Aug 1, 2019 | SentenceSentence segmentation | —Unverified | 0 | 0 |
| A Statistical, Grammar-Based Approach to Microplanning | Apr 1, 2017 | SentenceSentence segmentation | —Unverified | 0 | 0 |
| Optical Character Recognition, Word Segmentation, Sentence Segmentation, and Information Extraction for Historical and Literature Texts in Classical Chinese | Sep 1, 2020 | Optical Character RecognitionOptical Character Recognition (OCR) | —Unverified | 0 | 0 |
| When Classical Chinese Meets Machine Learning: Explaining the Relative Performances of Word and Sentence Segmentation Tasks | Jul 22, 2020 | BIG-bench Machine LearningSegmentation | —Unverified | 0 | 0 |
| Relative Positional Encoding for Speech Recognition and Direct Translation | May 20, 2020 | PositionSentence | —Unverified | 0 | 0 |
| Universal Joint Morph-Syntactic Processing: The Open University of Israel's Submission to The CoNLL 2017 Shared Task | Aug 1, 2017 | MORPHSentence | —Unverified | 0 | 0 |
| Alibaba Submission to the WMT20 Parallel Corpus Filtering Task | Nov 1, 2020 | DiversityLanguage Identification | —Unverified | 0 | 0 |
| Segmentation en phrases : ouvrez les guillemets sans perdre le fil | Jul 29, 2024 | SentenceSentence segmentation | —Unverified | 0 | 0 |
| Semi-supervised Thai Sentence Segmentation Using Local and Distant Word Representations | Aug 4, 2019 | SentenceSentence segmentation | —Unverified | 0 | 0 |
| Sentence Identification with BOS and EOS Label Combinations | Jan 31, 2023 | SentenceSentence segmentation | —Unverified | 0 | 0 |
| Sentence Segmentation for Classical Chinese Based on LSTM with Radical Embedding | Oct 5, 2018 | SegmentationSentence | —Unverified | 0 | 0 |
| Sentence Segmentation in Narrative Transcripts from Neuropsychological Tests using Recurrent Convolutional Neural Networks | Oct 2, 2016 | POSSegmentation | —Unverified | 0 | 0 |