Linguistic sequence complexity

id: linguistic-sequence-complexity-311-17234897
title: Linguistic sequence complexity
text: Linguistic sequence complexity (LC) is a measure of the 'vocabulary richness' of a genetic text in gene sequences. When a nucleotide sequence is written as text using a four-letter alphabet, the repetitiveness of the text, that is, the repetition of its N-grams (words), can be calculated and serves as a measure of sequence complexity. Thus, the more complex a DNA sequence, the richer its oligonucleotide vocabulary, whereas repetitious sequences have relatively lower complexities. Subsequent work
brand slug: wiki
category slug: encyclopedia
description:
original url: https://en.wikipedia.org/wiki/Linguistic_sequence_complexity
date created:
date modified: 2023-08-18T08:47:31Z
main entity: {"identifier":"Q6554066","url":"https://www.wikidata.org/entity/Q6554066"}
image:
fields total: 13
integrity: 13

Related Entries

Explore Next Part