## [Open Reading Frames](https://rosalind.info/problems/orf/)

### Background
Three immediate wrinkles of complexity arise when we try to pass directly from DNA to proteins.

First, not all DNA will be transcribed into RNA: so-called junk DNA appears to have no practical purpose for cellular function. 

Second, we can begin translation at any position along a strand of RNA, meaning that any substring of a DNA string can serve as a template for translation, as long as it begins with a start codon, ends with a stop codon, and has no other stop codons in the middle.

As a result, the same RNA string can actually be translated in three different ways, depending on how we group triplets of symbols into codons. For example, ...AUGCUGAC... can be translated as ...AUGCUG..., ...UGCUGA..., and ...GCUGAC..., which will typically produce wildly different protein strings.

### Open Reading Frames
Either strand of a DNA double helix can serve as the coding strand for RNA transcription. Hence, a given DNA string implies six total reading frames, or ways in which the same region of DNA can be translated into amino acids: three reading frames result from reading the string itself, whereas three more result from reading its reverse complement.

An open reading frame (ORF) is one which starts from the start codon and ends by stop codon, without any other stop codons in between. Thus, a candidate protein string is derived by translating an open reading frame into amino acids until a stop codon is reached.

### Problem
**Given:** A DNA string `s` of length at most 1 kbp in FASTA format.

**Return:** Every distinct candidate protein string that can be translated from ORFs of `s`. Strings can be returned in any order.

### Example
Input:
```
>Rosalind_99
AGCCATGTAGCTAACTCAGGTTACATGGGGATGACCCCGCGACTTGGATTAGAGTCTCTTTTGGAATAAGCCTGAATGATCCGAGTAGCATCTCAG
```

Output:
```
MLLGSFRLIPKETLIQVAGSSPCNLS
M
MGMTPRLGLESLLE
MTPRLGLESLLE
```

In [1]:
def protein_candidates(dna_strand: str) -> list:
    
    """
    Returns a list of every possible protein that may be
    translated from the open reading frames in the given strand.

    Args:
        dna_strand (str): DNA string of length at most 1 kbp.

    Returns:
        list: Every distinct candidate protein string that can be 
        translated from open reading frames in the given DNA strand.
    """
    
    candidates = []
    return candidates

In [2]:
import ipytest
ipytest.autoconfig()

def test_case_1():
    dna = "AGCCATGTAGCTAACTCAGGTTACATGGGGATGACCCCGCGACTTGG" + \
    "ATTAGAGTCTCTTTTGGAATAAGCCTGAATGATCCGAGTAGCATCTCAG"
    actual = protein_candidates(dna)
    expected = ["MLLGSFRLIPKETLIQVAGSSPCNLS", "M", 
        "MGMTPRLGLESLLE", "MTPRLGLESLLE"]
    assert actual == expected

ipytest.run()

[31mF[0m[31m                                                                                            [100%][0m
[31m[1m___________________________________________ test_case_1 ____________________________________________[0m

    [94mdef[39;49;00m [92mtest_case_1[39;49;00m():[90m[39;49;00m
        dna = [33m"[39;49;00m[33mAGCCATGTAGCTAACTCAGGTTACATGGGGATGACCCCGCGACTTGG[39;49;00m[33m"[39;49;00m + \
        [33m"[39;49;00m[33mATTAGAGTCTCTTTTGGAATAAGCCTGAATGATCCGAGTAGCATCTCAG[39;49;00m[33m"[39;49;00m[90m[39;49;00m
        actual = protein_candidates(dna)[90m[39;49;00m
        expected = [[33m"[39;49;00m[33mMLLGSFRLIPKETLIQVAGSSPCNLS[39;49;00m[33m"[39;49;00m, [33m"[39;49;00m[33mM[39;49;00m[33m"[39;49;00m,[90m[39;49;00m
            [33m"[39;49;00m[33mMGMTPRLGLESLLE[39;49;00m[33m"[39;49;00m, [33m"[39;49;00m[33mMTPRLGLESLLE[39;49;00m[33m"[39;49;00m][90m[39;49;00m
>       [94massert[39;49;00m actual == expected[90m[39;49;00m
[1m[31



<ExitCode.TESTS_FAILED: 1>