Space-Efficient Representation of Entity-centric Query Language Models

Christophe Van Gysel; Mirko Hannemann; Ernest Pusateri; Youssef; Oualil; Ilya Oparin

arXiv:2206.14885·cs.CL·July 1, 2022·1 cites

Space-Efficient Representation of Entity-centric Query Language Models

Christophe Van Gysel, Mirko Hannemann, Ernest Pusateri, Youssef, Oualil, Ilya Oparin

PDF

Open Access

TL;DR

This paper presents a space-efficient method for representing entity-centric query language models using probabilistic grammars within the FST framework, improving recognition accuracy for virtual assistants on resource-constrained devices.

Contribution

It introduces a deterministic approximation to probabilistic grammars that reduces resource usage and integrates with FSTs, enhancing on-device spoken entity recognition.

Findings

01

10% relative word error rate improvement on long tail entity queries

02

Efficient space representation of entity-centric language models

03

Complementary to n-gram models in ASR systems

Abstract

Virtual assistants make use of automatic speech recognition (ASR) to help users answer entity-centric queries. However, spoken entity recognition is a difficult problem, due to the large number of frequently-changing named entities. In addition, resources available for recognition are constrained when ASR is performed on-device. In this work, we investigate the use of probabilistic grammars as language models within the finite-state transducer (FST) framework. We introduce a deterministic approximation to probabilistic grammars that avoids the explicit expansion of non-terminals at model creation time, integrates directly with the FST framework, and is complementary to n-gram models. We obtain a 10% relative word error rate improvement on long tail entity queries compared to when a similarly-sized n-gram model is used without our method.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsTopic Modeling · Speech and dialogue systems · Natural Language Processing Techniques