Linguistic sequence complexity
id:
linguistic-sequence-complexity-311-17234897
title:
Linguistic sequence complexity
text:
Linguistic sequence complexity (LC) is a measure of the 'vocabulary richness' of a genetic text in gene sequences.
When a nucleotide sequence is written as text using a four-letter alphabet, the repetitiveness of the text, that is, the repetition of its N-grams (words), can be calculated and serves as a measure of sequence complexity. Thus, the more complex a DNA sequence, the richer its oligonucleotide vocabulary, whereas repetitious sequences have relatively lower complexities. Subsequent work
brand slug:
wiki
category slug:
encyclopedia
description:
original url:
https://en.wikipedia.org/wiki/Linguistic_sequence_complexity
date created:
date modified:
2023-08-18T08:47:31Z
main entity:
{"identifier":"Q6554066","url":"https://www.wikidata.org/entity/Q6554066"}
image:
fields total:
13
integrity:
13