• Home
  • Search
  • Structured Search in Annotated Document Collections
  • Cite Icon2
  • https://doi.org/10.1145/3289600.3290618Copy DOI Icon

Structured Search in Annotated Document Collections

  • Jan 30, 2019
  • Dhruv Gupta +1 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

In this work, we demonstrate structured search capabilities of the GYANI indexing infrastructure. GYANI allows linguists, journalists, and scholars in humanities to search large semantically annotated document collections in a structured manner by supporting queries with regular expressions between word sequences and annotations. In addition to this, we provide support for attaching semantics to words via annotations in the form of part-of-speech, named entities, temporal expressions, and numerical quantities. We demonstrate that by enabling such structured search capabilities we can quickly gather annotated text regions for various knowledge-centric tasks such as information extraction and question answering.

Similar Papers
  • Conference Article
  • Citations6

GYANI

  • Oct 17, 2018
  • Dhruv Gupta +1
  • Conference Article
  • Citations15

Extracting Events and Temporal Expressions from Text

  • Sep 01, 2010
  • Naushad Uzzaman +1
  • PDF
  • Research Article

Extracting Reproductive Condition and Habitat Information from Text Using a Transformer-based Information Extraction Pipeline

  • Sep 11, 2023
  • Biodiversity Information Science and Standards
  • Roselyn Gabud +3
  • Conference Article
  • Citations1

Identifying temporal expression and its syntactic role using FST and lexical data from corpus

  • Jan 01, 2000
  • Juntae Yoon +2
  • Research Article

Cyberformalism: Histories of Linguistic Forms in the Digital Archive by Daniel Shore

  • Jan 01, 2020
  • Modern Language Review
  • Yann Ciarán Ryan
  • Research Article
  • Citations53

An Information-Extraction Approach to Speech Processing: Analysis, Detection, Verification, and Recognition

  • May 01, 2013
  • Proceedings of the IEEE
  • Chin-Hui Lee +1
  • Book Chapter
  • Citations4

Information Extraction from Medical Texts with BERT Using Human-in-the-Loop Labeling

  • May 18, 2023
  • Hendrik Šuvalov +2
  • Book Chapter
  • Citations12

H $\imath$ LεX: A System for Semantic Information Extraction from Web Documents

  • Jan 01, 2008
  • Massimo Ruffolo +1
  • Dissertation

Language modeling approaches to question answering

  • Jul 01, 2009
  • Protima Banerjee +1
  • Single Book
  • Citations3

Memory-Based Parsing

  • Oct 31, 2004
  • Sandra Kübler
  • PDF
  • Conference Article
  • Citations4

Challenges for Information Extraction from Dialogue in Criminal Law

  • Jan 01, 2021
  • Jenny Hong +2
  • Research Article

Automated Taxonomy Induction and its Applications

  • Jan 01, 2017
  • Infoscience (Ecole Polytechnique Fédérale de Lausanne)
  • Amit Gupta
  • Conference Article

KnowledgeHub: An End-to-End Tool for Assisted Scientific Discovery

  • Aug 01, 2024
  • Yutao Sun +10
  • Book Chapter
  • Citations1

A Model for Information Extraction in Portuguese Based on Text Patterns

  • Jan 01, 2013
  • Tiago Luis Bonamigo +1
  • Book Chapter
  • Citations21

A Machine Learning Approach to Information Extraction

  • Jan 01, 2005
  • Alberto Téllez-Valero +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.