Nuc Seq Record
A NucSeqRecord consists of a NucSeq and several optional annotations.
Main attributes:
id - Identifier such as a locus tag (string)
seq - The sequence itself (NucSeq) Additional attributes:
name - Sequence name, e.g. gene name (string)
description - Additional text (string)
annotations - A map of strings containing key-value pairs of annotations for different features
letterAnnotations - Per letter/symbol annotation. This holds an ImmutableList whose length matches that of the sequence. A typical use would be to hold a list of integers representing sequencing quality scores.
from Bio.Seq import NucSeq from Bio.Seq import NucSeqRecord val record_1 = NucSeqRecord(NucSeq("ATCG"), "1", "seq1", "the first sequence")
Constructors
Functions
Returns the complement sequence of DNA or RNA, ambiguity is preserved
Return a new SeqRecord with the complement sequence. The sequence will have id id. If the other parameters are not specified, they will be taken from the original SeqRecord.
Returns the NUC of this NucSeq at the specified index i with i starting at zero Negative index start from the end of the sequence, i.e. -1 is the last base
Returns a subset NucSeq based on the IntRange of Nucleotides. Kotlin range operator is "..". Indices start at zero. Note Kotlin IntRange are inclusive end, while Python slices exclusive end Negative slices "-3..-1" start from the last base (i.e. would return the last three bases)
Returns a subset NucSeq based on the inclusive start i and inclusive last j Indices start at zero. Negative index start from the end of the sequence, i.e. -1 is the last base
Extracts the kmer count from a nucleotide sequence NucSeq. In general kmers from both strands should be extracted, unless strand specificity is really known. No kmers that include an ambiguous base pair will be included Using stepSize subsets of the kmers can be sampled
Returns the complement sequence of DNA or RNA
Return a new SeqRecord with the reverse complement sequence. The sequence will have id id. If the other parameters are not specified, they will be taken from the original SeqRecord.
Returns a string summary of the NucSeqRecord. Uses the representational string version of the sequence.
Translate a nucleotide sequence NucSeq into amino acids ProteinSeq.