Full interrogation of nuclease dsbs and sequencing (find-seq)
15 claims: 2 independent, 13 dependent
- 1A method of preparing a library of covalently closed DNA fragments, the method comprising:providing DNA, preferably genomic DNA (gDNA) from a cell type or organism of interest;randomly shearing the DNA to a defined average length to provide a population of DNA fragments;preparing the fragments for end-ligation,;ligating to the ends of the fragments a first hairpin adapter comprising at least a single deoxyuridine and a first primer site to prepare a population of ligated fragments;and purifying the ligated fragments using an exonuclease, thereby preparing a library of covalently closed DNA fragments.
- 5A method of preparing a library of fragments comprising nuclease-induced double stranded breaks in DNA, the method comprising:providing DNA;randomly shearing the DNA to a defined average length, preferably an average length of about 200-1000 base pairs (bps);end-repairing and then A-tailing the sheared DNA;ligating to the DNA a first hairpin adapter, comprising a first region, of about 10-20, nucleotides;a second region of about 45-65 nucleotides that forms one or more hairpin loops and comprises a first primer site compatible for use in PCR priming and/or sequencing;and a third region of about 10-20 nucleotides that is complementary to the first region, with a single deoxyuridine nucleotide between the first and second regions;contacting the sample with one or more exonucleases, sufficient to degrade any DNA molecules that lack the first hairpin adapter ligated to both of their ends;treating the sample with a nuclease to induce site-specific cleavage;end-repairing and then A-tailing the resulting ends;and ligating a second hairpin adapter comprising a first region of about 10-20 nucleotides;a second region of about 40-60 nucleotides that forms one or more hairpin loops and comprises a second primer compatible for use with the first primer site in PCR priming and/or sequencing;and a third region of about 10-20 nucleotides that is complementary to the first region and that also contains a single deoxyuridine nucleotide between the second and third regions, to create a population wherein the DNA fragments that were cleaved by the nuclease have a first and second hairpin adapter ligated to their respective ends;thereby preparing a library of fragments wherein one end was created by a nuclease-induced double stranded break in the DNA.
- 15The method of claims 1-14, wherein:the defined average length of the randomly sheared DNA is about 200-1000 base pairs (bps);the step of preparing the fragments for end ligation comprises end-reparing and then A-tailing the sheared DNA;the first hairpin adapter comprises a first region of about 10-20 nucleotides;a second region of about 45-65 nucleotides that forms one or more hairpin loops and the first primer site is compatible for use in PCR priming and/or sequencing;and a third region of about 10-20 nucleotides that is complementary to the first region, with a single deoxyuridine nucleotide between the first and second regions;and/or the second hairpin adapter comprising a first region of about 10-20 nucleotides;a second region of about 40-60 nucleotides that forms one or more hairpin loops and comprises a second primer compatible for use with the first primer site in PCR priming and/or sequencing;and a third region of about 10-20 nucleotides that is complementary to the first region and that also contains a single deoxyuridine nucleotide between the second and third regions.
Independent claims5
77 paragraphs in 8 sections, as filed
TECHNICAL FIELD
0001Described herein are <i>in vitro</i> methods for defining the genome-wide cleavage specificities of engineered nucleases such as CRISPR-Cas9 Nucleases.
BACKGROUND
0002Engineered nuclease technology including zinc fingers, TALENs, and CRISPR-Cas9 nucleases, is revolutionizing biomedical research and providing important new modalities for therapy of gene-based diseases.
SUMMARY
0003The invention is defined in the claims. At least in part, the present invention is based on the development of sensitive, unbiased methods for genome-wide detection of potential engineered nuclease (e.g., CRISPR-Cas9) off-target cleavage sites from cell type-specific genomic DNA samples. The present methods use exonuclease selection of covalently closed DNA molecules to create a population of genomic DNA molecules with very few free DNA ends, as a starting population for cleavage-specific enrichment and sequencing. Enrichment of these cleaved fragments, estimated to be >20,000X from human genomic DNA, enables very sequencing-efficient discovery of in vitro cleaved DNA fragments, in contrast to methods such as Digenome-Seq (<nplcit id="ncit0001" npl-type="s"><text>Kim et al., Nat Methods. 2015 Mar;12(3):237-43</text></nplcit>) that rely on whole-genome sequencing. After optimization, the present <i>in vitro</i> assay detected 100% of off-target cleavage sites detected by the in-cell GUIDE-seq assay at the VEGFA site target site used as a test case and described in <nplcit id="ncit0002" npl-type="s"><text>Tsai et al., Nat Biotechnol. 2015 Feb;33(2):187-97</text></nplcit> (in other words, the present in vitro assay detects a superset of GUIDE-seq detected cleavage sites).
0004Described herein are methods of enzymatically preparing a library of DNA fragments without DNA double-stranded breaks (DSBs) (the terms "breaks" and "cleavage" are used herein interchangeably) by ligation of stem-loop or hairpin adapters followed by exonuclease selection, and for detecting nuclease-induced cleavage of this enzymatically purified library of covalently closed DNA fragments by sequencing. Together, these two methods comprise a strategy for efficiently mining for nuclease-induced cleavage sites in complex mixtures of DNA. Thus the methods can include creating a population of molecules without ends, treating that population with a nuclease, and finding molecules in this population that have newly created ends as a result of nuclease-induced cleavage.
0005In a first aspect, the invention provides methods for preparing a library of covalently closed DNA fragments. The methods include providing DNA, e.g., gDNA from a cell type or organism of interest; randomly shearing the DNA to a defined average length, e.g., an average length of about 200-1000 bps, e.g., about 500 bps, to provide a population of DNA fragments; preparing the fragments for end-ligation, e.g., by end-repairing and then A-tailing the sheared DNA; ligating a first hairpin adapter comprising at least a single deoxyuridine (uracil) and a first primer site compatible for use in PCR priming and/or sequencing, e.g., next generation sequencing (NGS) to the ends of the fragments, to prepare a population of ligated fragments; and purifying the ligated fragments using an exonuclease, thereby preparing a library of covalently closed DNA fragments.
0006The methods can also include contacting the library of covalently closed DNA fragments with a nuclease to induce site-specific cleavage; ligating a second hairpin adapter comprising at least a single deoxyuridine and a second primer site compatible for use with the first primer site in PCR priming and/or sequencing; contacting the library with an enzyme, e.g., uracil DNA glycosylase (UDG) and/or endonuclease VIII, a DNA glycosylase-lyase, to nick the DNA at the deoxyuridine; and sequencing those fragments having a first and second hairpin adapter.
0007Also provided herein are methods for preparing a library of fragments comprising nuclease-induced double stranded breaks in DNA, e.g., genomic DNA (gDNA). The methods can include providing DNA, e.g., gDNA from a cell type or organism of interest; randomly shearing the DNA to a defined average length, e.g., an average length of about 200-1000 bps, e.g., about 100-500 bps; end-repairing and then A-tailing the sheared DNA; ligating a first single-tailed hairpin adapter, comprising a first region, of about 10-20, e.g., 12 nucleotides; a second region, of about 45-65, e.g., 58 nucleotides, that forms one or more hairpin loops and comprises a first primer site compatible for use in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region, of about 10-20, e.g., 13 nucleotides (e.g., one longer than the first region) that is complementary to the first region, with a single deoxyuridine nucleotide between the first and second regions; contacting the sample with one or more exonucleases (e.g., bacteriophage lambda exonuclease, E. coli ExoI, PlasmidSafe™ ATP-dependent exonuclease), sufficient to degrade any DNA molecules that lack the first (e.g., 5') single-tailed hairpin adapter ligated to both of their ends; treating the sample with a nuclease to induce site-specific cleavage (e.g., of on- and/or off-target sites, e.g., to induce blunt or staggered/overhanging ends) end-repairing and then A-tailing the resulting ends; ligating a second (e.g., 3') single-tailed hairpin adapter comprising a first region of about 10-20, e.g., 15 nucleotides; a second region of about 40-60, e.g., 48, nucleotides that forms one or more hairpin loops and comprises a second primer compatible for use with the first primer site in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region of about 10-20, e.g., 16 nucleotides (e.g., one longer than the first region) that is complementary to the first region and that also contains a single deoxyuridine nucleotide between the second and third regions, to create a population wherein the DNA fragments that were cleaved by the nuclease have a first and second single-tailed hairpin adapter ligated to their respective ends; thereby preparing a library of fragments, e.g., wherein one end was created by a nuclease-induced double stranded break in the DNA.
0008In some embodiments, the methods include contacting the library with uracil DNA glycosylase (UDG) and/or endonuclease VIII, a DNA glycosylase-lyase to nick the DNA at the deoxyuridine; and sequencing those fragments bearing a first and a second hairpin adapter.
0009Also provided herein are methods for detecting nuclease-induced double stranded breaks (DSBs) in DNA, e.g., in genomic DNA (gDNA) of a cell. The methods include providing DNA, e.g., gDNA from a cell type or organism of interest; randomly shearing the DNA to a defined average length, e.g., an average length of about 200-1000 bps, e.g., about 500 bps; end-repairing and then A-tailing the sheared DNA; ligating a first single-tailed hairpin adapter, preferably comprising a first region, e.g., of about 10-20, e.g., 12 nucleotides; a second region, e.g., of about 45-65, e.g., 58 nucleotides, that forms one or more hairpin loops and comprises a first primer site compatible for use in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region, e.g., of about 10-20, e.g., 13 nucleotides (e.g., one longer than the first region) that is complementary to the first region, with a single deoxyuridine nucleotide between the first and second regions; contacting the sample with one or more exonucleases (e.g., bacteriophage lambda exonuclease, E. coli ExoI, PlasmidSafe™ ATP-dependent exonuclease), sufficient to degrade any DNA molecules that lack the first single-tailed hairpin adapter ligated to both of their ends; treating the sample with a nuclease to induce site-specific cleavage the DNA, e.g., to produse on- and/or off-target cleavage sites; optionally end-repairing and then A-tailing the resulting cleaved ends; ligating a second (e.g., 3') single-tailed hairpin adapter comprising a first region of about 10-20, e.g., 15 nucleotides; a second region of about 40-60, e.g., 48, nucleotides that forms one or more hairpin loops and comprises a second primer compatible for use with the first primer site in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region of about 10-20, e.g., 16 nucleotides (e.g., one longer than the first region) that is complementary to the first region, and that also contains a single deoxyuridine nucleotide between the second and third regions, to create a population wherein the DNA fragments that were cleaved by the nuclease have the first and second hairpin adapters ligated to their respective ends; thereby preparing a library of fragments comprising nuclease-induced double stranded breaks in the DNA, e.g., wherein one end was created by a nuclease-induced double stranded break in the DNA; contacting the library with uracil DNA glycosylase (UDG) and/or endonuclease VIII, a DNA glycosylase-lyase to nick the DNA at the deoxyuridine; and sequencing those fragments bearing a first and a second hairpin adapter; thereby detecting DSBs induced by the nuclease.
0010In some embodiments, the engineered nuclease is selected from the group consisting of meganucleases, MegaTALs, zinc-finger nucleases, transcription activator effector-like nucleases (TALEN), and Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR)/Cas RNA-guided nucleases (CRISPR/Cas RGNs).
0011In some embodiments, treating the sample with a nuclease to induce site-specific cleavage, e.g., at on- and off-target sites, comprises contacting the sample with a Cas9 nuclease complexed with a specific guide RNA (gRNA).
0012Further, provided herein are methods for determining which of a plurality of guide RNAs is most specific, i.e., induces the fewest off-target DSBs. The methods include, for each of the plurality of guide RNAs: providing gDNA from a cell type or organism of interest; randomly shearing the gDNA to a defined average length, e.g., an average length of about 200-1000 bps, e.g., about 500 bps; end-repairing and then A-tailing the sheared gDNA; ligating a first single-tailed hairpin adapter, preferably comprising a first region, e.g., of about 10-20, e.g., 12 nucleotides; a second region, e.g., of about 45-65, e.g., 58 nucleotides, that forms one or more hairpin loops and comprises a first primer site compatible for use in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region, e.g., of about 10-20, e.g., 13 nucleotides (e.g., one longer than the first region) that is complementary to the first region, with a single deoxyuridine nucleotide between the first and second regions; contacting the sample with one or more exonucleases (e.g., bacteriophage lambda exonuclease, E. coli ExoI, PlasmidSafe™ ATP-dependent exonuclease), sufficient to degrade any DNA molecules that lack the first hairpin adapter ligated to both of their ends; treating the sample with a Cas9 nuclease compatible with the guide RNA to induce site-specific cleavage (e.g., of on- and/or off-target sites, e.g., to produce blunt or staggered/overhanging ends); optionally end-repairing and then A-tailing the resulting cleaved ends; ligating a second (e.g., 3') single-tailed hairpin adapter comprising a first region of about 10-20, e.g., 15 nucleotides; a second region of about 40-60, e.g., 48, nucleotides that forms one or more hairpin loops and comprises a second primer compatible for use with the first primer site in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region of about 10-20, e.g., 16 nucleotides (e.g., one longer than the first region) that is complementary to the first region, and that also contains a single deoxyuridine nucleotide between the second and third regions, to create a population wherein the DNA fragments that have a first and second hairpin adapter ligated to their ends are those that were cleaved by the nuclease; thereby preparing a library of fragments comprising nuclease-induced double stranded breaks in DNA, e.g., wherein one end was created by a nuclease-induced double stranded break in the DNA; contacting the library with uracil DNA glycosylase (UDG) and/or endonuclease VIII, a DNA glycosylase-lyase to nick the DNA at the deoxyuridine; and sequencing those fragments bearing a first and a second hairpin adapter, thereby detecting DSBs induced by the nuclease in each sample; optionally identifying whether each DSB is on-target or off-target; comparing the DSBs induced by the nuclease in each sample; and determining which of the plurality of guide RNAs induced the fewest off-target DSBs.
0013In some embodiments, the DNA is isolated from a mammalian, plant, bacterial, or fungal cell (e.g., gDNA).
0014In some embodiments, the DNA is synthetic.
0015In some embodiments, the engineered nuclease is a TALEN, zinc finger, meganuclease, megaTAL, or a Cas9 nuclease.
0016In some embodiments, the engineered nuclease is a Cas9 nuclease, and the method also includes expressing in the cells a guide RNA that directs the Cas9 nuclease to a target sequence in the genome.
0017In some embodiments, the primer site in the first or second hairpin adapter comprises a next generation sequencing primer site, a randomized DNA barcode or unique molecular identifier (UMI).
0018The present methods have several advantages. For example, the present methods are <i>in vitro</i>; in contrast, GUIDE-seq is cell-based, requiring the introduction of double stranded oligodeoxynucleotides (dsODN) as well as expression or introduction of nuclease or nuclease-encoding components into cells. Not all cells will allow the introduction of dsODNs and/or nuclease or nuclease-encoding components, and these reagents can be toxic in some cases. If a cell type of particular interest is not amenable to introduction of dsODNs, nucleases, or nuclease-encoding components, a surrogate cell type might be used, but cell-specific effects might not be detected. The present methods do not require delivery of the dsODN, nuclease, or nuclease-encoding components into a cell, are not influenced by chromatin state, do not create toxicity issues, and enable interrogation of specific cellular genomes by cleavage of genomic DNA that is obtained from the cell-type of interest.
0019Unless otherwise defined, all technical and scientific terms used herein have the same meaning as commonly understood by one of ordinary skill in the art to which this invention belongs. Methods and materials are described herein for use in the present invention; other, suitable methods and materials known in the art can also be used. The materials, methods, and examples are illustrative only and not intended to be limiting. In case of conflict, the present specification, including definitions, will control.
0020Other features and advantages of the invention will be apparent from the following detailed description and figures, and from the claims.
DESCRIPTION OF DRAWINGS
0021<ul id="ul0001" list-style="none" compact="compact"><li><figref idref="f0001 f0002">Figures 1A-B</figref>. Overview of an exemplary in vitro assay for genome-wide identification of CRISPR/Cas9 cleavage sites from complex mixtures of DNA called FIND-seq (Full Interrogation of Nuclease DSBs). (a) Genomic DNA is isolated from human cells and sheared to an average of ∼500 bp using a Covaris S220 AFA instrument. Sheared DNA is end-repaired, A-tailed, and ligated with a 5' single-tailed hairpin adapter containing a single deoxyuradine. Successful ligation of the hairpin adapter to both ends of the DNA fragments will result in covalently closed DNA with no free end. Lambda Exonuclease and E. coli ExoI or PlasmidSafe ATP-dependent exonuclease can then be used to dramatically reduce the background of DNA with free ends, leaving a predominantly uniform population of covalently closed DNA molecules. Cas9 cleavage of adapter-ligated gDNA will produce ends with free 5' phosphates for the ligation of a second 3' single-tailed hairpin adapter that also contains a single uracil base. Treatment with USER enzyme mixture and PCR enriches for DNA molecules that have been ligated with both 5' and 3' single-tailed adapters. The use of single-tailed adapters reduces background, as only molecules that contain both 5' and 3' adapters can be amplified and sequenced. (b) Detailed schematic of adapter ligation steps, with sequence and predicted DNA secondary structure of hairpin adapters. Ligation of hairpin adapters at both ends creates covalently closed DNA molecules that can be subsequently reopened at uracil-containing sites.</li><li><figref idref="f0003 f0004">Figures 2A-B</figref>. Reads mapped at on-target and off-target cleavage sites. (a) Visualization of 'on-target' genomic regions where bidirectionally mapping uniform-end read signatures of Cas9 cleavage can be identified. Reads are indicated by rectangles; coverage is displayed above. The arrow indicates the predicted cut site and data is shown for 3 target sites in the VEGFA locus. A sample treated with <i>Streptococcus pyogenes</i> Cas9 complexed with EMX1 gRNA is also shown as a negative control. (b) Visualization of example 'off-target' genomic region 20:56175349-56175372 in VEGFA site 1. Bidirectionally mapping reads with 'uniform' ends can be observed in this region and are typical of sites detected by this in vitro assay. Reads generally start within 1-bp of the predicted cut-site and originate outward in both directions from this position.</li><li><figref idref="f0005 f0006">Figures 3A-D</figref>. Analysis of FIND-seq in vitro cleavage selection assay results using different exonuclease treatment conditions. (a) Mapped read counts at genome-wide in vitro cleavage assay detected sites for a under different exonuclease treatments. (b) Number of sites detected by genome-wide in vitro cleavage assay with different exonuclease treatments. (c) Average read counts at genome-wide in vitro cleavage assay detected sites under different exonuclease treatments. (d) Percentage of GUIDE-seq detected sites that are detected by genome-wide in vitro cleavage assay. Note there are three conditions where 100% of GUIDE-seq sites are detected using the simple in vitro assay.</li><li><figref idref="f0007">Figure 4</figref>. Venn diagram of overlap between GUIDE-seq and FIND-seq genomic in vitro cleavage assay performed with 1 hour of lambda exonuclease and E. coli exonuclease I, and 1 hour of PlasmidSafe exonuclease treatment at the <i>VEGFA</i> target site 1. In this condition, 100% of GUIDE-seq sites are also detected with the in vitro assay.</li><li><figref idref="f0008">Figures 5A-B</figref>. Analysis of FIND-seq genomic in vitro cleavage assay testing the adapter ligation order with different exonuclease treatment conditions. The condition that detected the highest number of sites was 1 hr PlasmidSafe, 1 hr Lambda exonuclease/E. Coli exonuclease I treatment, with the 3' adapter added first. (a) Number of sites by FIND-seq genome-wide in vitro cleavage assay with different exonuclease treatments and adapter ligation order. (b) Read counts at FIND-seq detected off-target sites using different exonuclease treatments and adapter ligation order.</li></ul>
DETAILED DESCRIPTION
0022The invention is defined in the claims. An important consideration for the therapeutic deployment of CRISPR-Cas9 nucleases is having robust, comprehensive, unbiased, and highly sensitive methods for defining their off-target effects. Recently, a number of methods for defining the genome-wide off-target effects of CRISPR-Cas9 and other customizable nucleases have been described. For example, the GUIDE-seq method, which relies on uptake of a short double-stranded oligonucleotide "tag" into nuclease-induced DSBs in living cells, has been shown to define off-target sites on a genome-wide scale, identifying sites that are mutagenized with frequencies as low as 0.1% of the time in a population of cells (<nplcit id="ncit0003" npl-type="s"><text>Tsai et al., Nat Biotechnol. 2015</text></nplcit>). Other cell-based methods for defining nuclease-induced off-target breaks include a method that maps translocation fusions to the on-target site and another that relies on uptake of integration-deficient lentivirus (IDLV) genomes into sites of DSBs.
0023Despite this recent progress, cell-based methods for off-target determination have a number of limitations including: (1) a requirement to be able to introduce both the nuclease components and a tag such as the dsODN or IDLV genome into cells; (2) biological selection pressures that might favor or disfavor the growth of cells harboring certain types of off-target mutations; and/or (3) the potential confounding effects of cell-type-specific parameters such as chromatin, DNA methylation, gene expression, and nuclear architecture on nuclease off-target activities/effects.
0024In vitro methods using purified genomic DNA provide an attractive alternative because they would sidestep these various limitations of cell-based approaches. However, in vitro methods face the challenge that isolated genomic DNA is by experimental necessity randomly sheared (or broken) into smaller pieces. This poses a challenge because it is not easy to differentially identify DSBs induced by shearing from those induced by treatment of the genomic DNA in vitro by nucleases. The recently described Digenome attempted to use alignment of common ends induced by nucleases in genomic sequence but the signal for this type of event can be challenging to discern relative to the background of random DSBs from shearing of genomic DNA induced during genomic DNA isolation and by deliberate shearing of the DNA to smaller size pieces required for DNA sequencing methods. In addition, and perhaps more importantly, because there is no enrichment for the nuclease-induced DSBs, nearly all of the sequencing data generated with Digenome is just re-sequencing of random genomic DNA, a factor that further limits the sensitivity of the method since very few reads generated contribute the desired information about nuclease-induced DSBs. Indeed, Digenome failed to identify a number of off-target sites found by GUIDE-seq for a particular gRNA, although the caveat must be added that the two experiments were performed in different cell lines.
0025Described herein is an <i>in vitro</i> method that enables comprehensive determination of DSBs induced by nucleases on any genomic DNA of interest. This method enables enrichment of nuclease-induced DSBs over random DSBs induced by shearing of genomic DNA. An overview of how the method works can be found in <figref idref="f0001"><b>Figures 1A</b></figref><b>and</b><figref idref="f0002"><b>1B</b></figref><b>.</b> In brief, genomic DNA from a cell type or organism of interest is randomly sheared to a defined average length. In the present experiments, an average length of 500 bps worked well for subsequent next-generation sequencing steps; however, shorter or longer lengths can also be used. The random broken ends of this genomic DNA are end repaired (i.e., the random overhanging ends are filled in/blunted, e.g., using T4 polymerase, Klenow fragment, and T4 Polynucleotide Kinase (PNK)) and then A-tailed (i.e., an A (adenosine) is enzymatically added to the 3' end of a blunt, double-stranded DNA molecule). This then enables the ligation of a hairpin adapter, preferably bearing a 5' end next-generation sequencing primer site, that also contains a single deoxyuridine nucleotide at a specific position <b>(</b><figref idref="f0002"><b>Figure 1B</b></figref><b>)</b> and a 1-nucleotide (thymidine) overhang at the 3' end; hereafter this adapter is referred to as the "5' single-tailed hairpin adapter" or simply "5' hairpin adapter".
0026Following ligation of this adapter, the sample is then treated with one or more exonucleases (e.g., bacteriophage lambda exonuclease, E. coli ExoI, PlasmidSafe™ ATP-dependent exonuclease) that degrade any DNA molecules that have not had the 5' single-tailed hairpin adapter ligated to both of their ends. The collection of molecules is then treated to induce blunt-end cuts, e.g., with a Cas9 nuclease complexed with a specific guide RNA (gRNA). The resulting nuclease-induced blunt ends are A-tailed and then a second hairpin adapter bearing a 3' end next-generation sequencing primer site and that also contains a single deoxyuridine nucleotide at a specific position is ligated to these ends; this adapter is also referred to herein as the "3' single-tailed hairpin adapter" or simply "3' hairpin adapter".
0027As a result of these treatments, the only DNA fragments that should have a 5' and a 3' single-tailed hairpin adapter ligated to their ends are those that were cleaved by the nuclease. Following treatment (e.g., by the USER enzyme mixture) to nick DNA wherever a deoxyuridine is present, only those fragments bearing the two types of adapters can be sequenced.
0028The present methods can be used to prepare libraries of fragments for next generation sequencing, e.g., to identify guideRNA/Cas9 combinations that induce the most specific DSBs (e.g., that have the fewest off-target effects), e.g., for therapeutic or research purposes.
Single-Tailed Hairpin Adapters
0029The present methods include the use of non-naturally occurring 3' and 5' single-tailed hairpin adapters. The hairpin adapters include (from 5' to 3') a first region, e.g., of about 10-20, e.g., 12 or 15, nucleotides; a second region, e.g., of about 45-65, e.g., 58 or 48, nucleotides that forms one or more hairpin loops and includes a sequence compatible for use in PCR priming and/or sequencing, e.g., next generation sequencing (NGS); and a third region, e.g., of about 10-20, e.g., 13 or 16, nucleotides that is complementary to the first region. The lengths of the first, second and third regions can vary depending on the NGS method selected, as they are dependent on the sequences that are necessary for priming for use with the selected NGS platform.
0030The hairpin adapters include at least one, preferably only one, uracil that allows the adaptor to be opened by Uracil DNA glycosylase (UDG) and Endonuclease VIII, a DNA glycosylase-lyase, e.g., the USER (Uracil-Specific Excision Reagent) Enzyme mixture (New England BioLabs). The UDG catalyzes the excision of uracil bases to form an abasic site but leave the phosphodiester backbone intact (see, e.g., <nplcit id="ncit0004" npl-type="s"><text>Lindhal et al., J. Biol. Chem.. 252:3286-3294 (1977</text></nplcit>); <nplcit id="ncit0005" npl-type="s"><text>Lindhal, Annu. Rev. Biochem. 51:61-64 (1982</text></nplcit>)). The Endonuclease VIII breaks the phosphodiester backbone at the 3' and 5' sides of the abasic site (see, e.g., <nplcit id="ncit0006" npl-type="s"><text>Melamede et al., Biochemistry 33:1255-1264 (1994</text></nplcit>); <nplcit id="ncit0007" npl-type="s"><text>Jiang et al., J. Biol. Chem. 272:32230-32239 (1997</text></nplcit>)). This combination generates a single nucleotide gap at the location of a uracil. In some embodiments, in the 5' single-tailed hairpin adapters the uracil is placed at or within 1, 2, 3, or 4 nucleotides of the end of the first region and the beginning of the second region, and in the 3' single-tailed hairpin adapters the uracil is placed at or within 1, 2, 3, or 4 nucleotides of the end of the first region and the beginning of the second region.
0031In the present methods, all parts of the hairpin adapters are preferably orthologous to the genome of the cell (i.e., are not present in or complementary to a sequence present in, i.e., have no more than 10%, 20%, 30%, 40%, or 50% identity to a sequence present in, the genome of the cell). The hairpin adapters can preferably be between 65 and 95 nts long, e.g., 70-90 nts or 75-85 nts long.
0032Each of the 5' and 3' single-tailed hairpin adapters should include a primer site that is a randomized DNA barcode (e.g., SHAPE-SEQ, <nplcit id="ncit0008" npl-type="s"><text>Lucks et al., Proc Natl Acad Sci U S A 108: 11063-11068</text></nplcit>), unique molecular identifier (UMI) (see, e.g., <nplcit id="ncit0009" npl-type="s"><text>Kivioja et al., Nature Methods 9, 72-74 (2012</text></nplcit>); <nplcit id="ncit0010" npl-type="s"><text>Islam et al., Nature Methods 11, 163-166 (2014</text></nplcit>); <nplcit id="ncit0011" npl-type="s"><text>Karlsson et al., Genomics. 2015 Mar;105(3):150-8</text></nplcit>), or unique PCR priming sequence and/or unique sequence compatible for use in sequencing (e.g., NGS). The sequence compatible for use in sequencing can be selected for use with a desired sequencing method, e.g., a next generation sequencing method, e.g., Illumina, Ion Torrent or library preparation method like Roche/454, Illumina Solexa Genome Analyzer, the Applied Biosystems SOLiD™ system, Ion Torrent™ semiconductor sequence analyzer, PacBio® real-time sequencing and Helicos™ Single Molecule Sequencing (SMS). See, e.g., <patcit id="pcit0001" dnum="WO2014020137A"><text>WO2014020137</text></patcit>, <nplcit id="ncit0012" npl-type="s"><text>Voelkerding et al., Clinical Chemistry 55:4 641-658 (2009</text></nplcit>) and <nplcit id="ncit0013" npl-type="s"><text>Metzker, Nature Reviews Genetics 11:31-46 (2010</text></nplcit>)). A number of kits are commercially available for preparing DNA for NGS, including the ThruPLEX DNA-seq Kit (Rubicon; see <patcit id="pcit0002" dnum="US7803550B"><text>U.S. Patents 7,803,550</text></patcit>; <patcit id="pcit0003" dnum="US8071312B"><text>8,071,312</text></patcit>; <patcit id="pcit0004" dnum="US8399199B"><text>8,399,199</text></patcit>; <patcit id="pcit0005" dnum="US8728737B"><text>8,728,737</text></patcit>) and NEBNext® (New England BioLabs; see e.g., <patcit id="pcit0006" dnum="US8420319B"><text>U.S. patent 8,420,319</text></patcit>)). Exemplary sequences are shown in Table A, below.
0033In some embodiments, the hairpin adapters include a restriction enzyme recognition site, preferably a site that is relatively uncommon in the genome of the cell.
0034The hairpin adapters are preferably modified; in some embodiments, the 5' ends of the hairpin adapters are phosphorylated. In some embodiments, the hairpin adapters are blunt ended. In some embodiments, the hairpin adapters include a random variety of 1, 2, 3, 4 or more nucleotide overhangs on the 5' or 3' ends, or include a single T at the 5' or 3' end.
0035The hairpin adapters can also include one or more additional modifications, e.g., as known in the art or described in <patcit id="pcit0007" dnum="US2011060493W" dnum-type="L"><text>PCT/US2011/060493</text></patcit>. For example, in some embodiments, the hairpin adapters is biotinylated. The biotin can be anywhere internal to the hairpin adapters (e.g., a modified thymidine residue (Biotin-dT) or using biotin azide), but not on the 5' or 3' ends. This provides an alternate method of recovering fragments that contain the FIND-seq hairpin adapters. Whereas in some embodiments, these sequences are retrieved and identified by PCR, in this approach they are physically pulled down and enriched by using the biotin, e.g., by binding to streptavidin-coated magnetic beads, or using solution hybrid capture; see, e.g., <nplcit id="ncit0014" npl-type="s"><text>Gnirke et al., Nature Biotechnology 27, 182 - 189 (2009</text></nplcit>).
0036Although the present working examples include ligating the 5' adapter and then the 3' adapter, the adapters can be added in either order, i.e., 5' adapter then 3' adapter, or 3' adapter then 5' adapter. The order may be optimized depending on the exonuclease treatment used, e.g., for Lambda exonuclease, which is a highly processive 5' -> 3' exonuclease, where adding the 5' adapter second may be advantageous.
Engineered Nucleases
0037There are presently four main classes of engineered nucleases: 1) meganucleases, 2) zinc-finger nucleases, 3) transcription activator effector-like nucleases (TALEN), and 4) Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR) Cas RNA-guided nucleases (RGN). See, e.g., <nplcit id="ncit0015" npl-type="s"><text>Gaj et al., Trends Biotechnol. 2013 Jul;31(7):397-405</text></nplcit>. Any of these, or variants thereof, can be used in the present methods. The nuclease can be transiently or stably expressed in the cell, using methods known in the art; typically, to obtain expression, a sequence encoding a protein is subcloned into an expression vector that contains a promoter to direct transcription. Suitable eukaryotic expression systems are well known in the art and described, e.g., in <nplcit id="ncit0016" npl-type="b"><text>Sambrook et al., Molecular Cloning, A Laboratory Manual (4th ed. 2013</text></nplcit>); <nplcit id="ncit0017" npl-type="b"><text>Kriegler, Gene Transfer and Expression: A Laboratory Manual (2006</text></nplcit>); and <nplcit id="ncit0018" npl-type="b"><text>Current Protocols in Molecular Biology (Ausubel et al., eds., 2010</text></nplcit>). Transformation of eukaryotic and prokaryotic cells are performed according to standard techniques (see, e.g., the reference above and <nplcit id="ncit0019" npl-type="s"><text>Morrison, 1977, J. Bacteriol. 132:349-351</text></nplcit>; <nplcit id="ncit0020" npl-type="b"><text>Clark-Curtiss & Curtiss, Methods in Enzymology 101:347-362 (Wu et al., eds, 1983</text></nplcit>).
Homing Meganucleases
0038Meganucleases are sequence-specific endonucleases originating from a variety of organisms such as bacteria, yeast, algae and plant organelles. Endogenous meganucleases have recognition sites of 12 to 30 base pairs; customized DNA binding sites with 18bp and 24bp-long meganuclease recognition sites have been described, and either can be used in the present methods and constructs. See, e.g., <nplcit id="ncit0021" npl-type="s"><text>Silva, G., et al., Current Gene Therapy, 11:11-27, (2011</text></nplcit>); <nplcit id="ncit0022" npl-type="s"><text>Arnould et al., Journal of Molecular Biology, 355:443-58 (2006</text></nplcit>); <nplcit id="ncit0023" npl-type="s"><text>Arnould et al., Protein Engineering Design & Selection, 24:27-31 (2011</text></nplcit>); and <nplcit id="ncit0024" npl-type="s"><text>Stoddard, Q. Rev. Biophys. 38, 49 (2005</text></nplcit>); <nplcit id="ncit0025" npl-type="s"><text>Grizot et al., Nucleic Acids Research, 38:2006-18 (2010</text></nplcit>).
CRISPR-Cas Nucleases
0039Recent work has demonstrated that clustered, regularly interspaced, short palindromic repeats (CRISPR)/CRISPR-associated (Cas) systems (<nplcit id="ncit0026" npl-type="s"><text>Wiedenheft et al., Nature 482, 331-338 (2012</text></nplcit>); <nplcit id="ncit0027" npl-type="s"><text>Horvath et al., Science 327, 167-170 (2010</text></nplcit>); <nplcit id="ncit0028" npl-type="s"><text>Terns et al., Curr Opin Microbiol 14, 321-327 (2011</text></nplcit>)) can serve as the basis of a simple and highly efficient method for performing genome editing in bacteria, yeast and human cells, as well as <i>in vivo</i> in whole organisms such as fruit flies, zebrafish and mice (<nplcit id="ncit0029" npl-type="s"><text>Wang et al., Cell 153, 910-918 (2013</text></nplcit>); <nplcit id="ncit0030" npl-type="s"><text>Shen et al., Cell Res (2013</text></nplcit>); <nplcit id="ncit0031" npl-type="s"><text>Dicarlo et al., Nucleic Acids Res (2013</text></nplcit>); <nplcit id="ncit0032" npl-type="s"><text>Jiang et al., Nat Biotechnol 31, 233-239 (2013</text></nplcit>); <nplcit id="ncit0033" npl-type="s"><text>Jinek et al., Elife 2, e00471 (2013</text></nplcit>); <nplcit id="ncit0034" npl-type="s"><text>Hwang et al., Nat Biotechnol 31, 227-229 (2013</text></nplcit>); <nplcit id="ncit0035" npl-type="s"><text>Cong et al., Science 339, 819-823 (2013</text></nplcit>); <nplcit id="ncit0036" npl-type="s"><text>Mali et al., Science 339, 823-826 (2013c</text></nplcit>); <nplcit id="ncit0037" npl-type="s"><text>Cho et al., Nat Biotechnol 31, 230-232 (2013</text></nplcit>); <nplcit id="ncit0038" npl-type="s"><text>Gratz et al., Genetics 194(4):1029-35 (2013</text></nplcit>)). The Cas9 nuclease from <i>S. pyogenes</i> can be guided via simple base pair complementarity between 17-20 nucleotides of an engineered guide RNA (gRNA), e.g., a single guide RNA or crRNA/tracrRNA pair, and the complementary strand of a target genomic DNA sequence of interest that lies next to a protospacer adjacent motif (PAM), e.g., a PAM matching the sequence NGG or NAG (<nplcit id="ncit0039" npl-type="s"><text>Shen et al., Cell Res (2013</text></nplcit>); <nplcit id="ncit0040" npl-type="s"><text>Dicarlo et al., Nucleic Acids Res (2013</text></nplcit>); <nplcit id="ncit0041" npl-type="s"><text>Jiang et al., Nat Biotechnol 31, 233-239 (2013</text></nplcit>); <nplcit id="ncit0042" npl-type="s"><text>Jinek et al., Elife 2, e00471 (2013</text></nplcit>); <nplcit id="ncit0043" npl-type="s"><text>Hwang et al., Nat Biotechnol 31, 227-229 (2013</text></nplcit>); <nplcit id="ncit0044" npl-type="s"><text>Cong et al., Science 339, 819-823 (2013</text></nplcit>); <nplcit id="ncit0045" npl-type="s"><text>Mali et al., Science 339, 823-826 (2013c</text></nplcit>); <nplcit id="ncit0046" npl-type="s"><text>Cho et al., Nat Biotechnol 31, 230-232 (2013</text></nplcit>); <nplcit id="ncit0047" npl-type="s"><text>Jinek et al., Science 337, 816-821 (2012</text></nplcit>)).
0040In some embodiments, the present system utilizes a wild type or variant Cas9 protein from <i>S. pyogenes</i> or <i>Staphylococcus aureus</i>, either as encoded in bacteria or codon-optimized for expression in mammalian cells. The guide RNA is expressed in the cell together with the Cas9. Either the guide RNA or the nuclease, or both, can be expressed transiently or stably in the cell.
TAL Effector Repeat Arrays
0041TAL effectors of plant pathogenic bacteria in the genus Xanthomonas play important roles in disease, or trigger defense, by binding host DNA and activating effector-specific host genes. Specificity depends on an effector-variable number of imperfect, typically ∼33-35 amino acid repeats. Polymorphisms are present primarily at repeat positions 12 and 13, which are referred to herein as the repeat variable-diresidue (RVD). The RVDs of TAL effectors correspond to the nucleotides in their target sites in a direct, linear fashion, one RVD to one nucleotide, with some degeneracy and no apparent context dependence. In some embodiments, the polymorphic region that grants nucleotide specificity may be expressed as a triresidue or triplet.
0042Each DNA binding repeat can include a RVD that determines recognition of a base pair in the target DNA sequence, wherein each DNA binding repeat is responsible for recognizing one base pair in the target DNA sequence. In some embodiments, the RVD can comprise one or more of: HA for recognizing C; ND for recognizing C; HI for recognizing C; HN for recognizing G; NA for recognizing G; SN for recognizing G or A; YG for recognizing T; and NK for recognizing G, and one or more of: HD for recognizing C; NG for recognizing T; NI for recognizing A; NN for recognizing G or A; NS for recognizing A or C or G or T; N* for recognizing C or T, wherein * represents a gap in the second position of the RVD; HG for recognizing T; H* for recognizing T, wherein * represents a gap in the second position of the RVD; and IG for recognizing T.
0043TALE proteins may be useful in research and biotechnology as targeted chimeric nucleases that can facilitate homologous recombination in genome engineering (e.g., to add or enhance traits useful for biofuels or biorenewables in plants). These proteins also may be useful as, for example, transcription factors, and especially for therapeutic applications requiring a very high level of specificity such as therapeutics against pathogens (e.g., viruses) as non-limiting examples.
0044Methods for generating engineered TALE arrays are known in the art, see, e.g., the fast ligation-based automatable solid-phase high-throughput (FLASH) system described in USSN <patcit id="pcit0008" dnum="US61610212B"><text>61/610,212</text></patcit>, and <nplcit id="ncit0048" npl-type="s"><text>Reyon et al., Nature Biotechnology 30,460-465 (2012</text></nplcit>); as well as the methods described in <nplcit id="ncit0049" npl-type="s"><text>Bogdanove & Voytas, Science 333, 1843-1846 (2011</text></nplcit>); <nplcit id="ncit0050" npl-type="s"><text>Bogdanove et al., Curr Opin Plant Biol 13, 394-401 (2010</text></nplcit>); <nplcit id="ncit0051" npl-type="s"><text>Scholze & Boch, J. Curr Opin Microbiol (2011</text></nplcit>); <nplcit id="ncit0052" npl-type="s"><text>Boch et al., Science 326, 1509-1512 (2009</text></nplcit>); <nplcit id="ncit0053" npl-type="s"><text>Moscou & Bogdanove, Science 326, 1501 (2009</text></nplcit>); <nplcit id="ncit0054" npl-type="s"><text>Miller et al., Nat Biotechnol 29, 143-148 (2011</text></nplcit>); <nplcit id="ncit0055" npl-type="s"><text>Morbitzer et al., T. Proc Natl Acad Sci U S A 107, 21617-21622 (2010</text></nplcit>); <nplcit id="ncit0056" npl-type="s"><text>Morbitzer et al., Nucleic Acids Res 39, 5790-5799 (2011</text></nplcit>); <nplcit id="ncit0057" npl-type="s"><text>Zhang et al., Nat Biotechnol 29, 149-153 (2011</text></nplcit>); <nplcit id="ncit0058" npl-type="s"><text>Geissler et al., PLoS ONE 6, e19509 (2011</text></nplcit>); <nplcit id="ncit0059" npl-type="s"><text>Weber et al., PLoS ONE 6, e19722 (2011</text></nplcit>); <nplcit id="ncit0060" npl-type="s"><text>Christian et al., Genetics 186, 757-761 (2010</text></nplcit>); <nplcit id="ncit0061" npl-type="s"><text>Li et al., Nucleic Acids Res 39, 359-372 (2011</text></nplcit>); <nplcit id="ncit0062" npl-type="s"><text>Mahfouz et al., Proc Natl Acad Sci U S A 108, 2623-2628 (2011</text></nplcit>); <nplcit id="ncit0063" npl-type="s"><text>Mussolino et al., Nucleic Acids Res (2011</text></nplcit>); <nplcit id="ncit0064" npl-type="s"><text>Li et al., Nucleic Acids Res 39, 6315-6325 (2011</text></nplcit>); <nplcit id="ncit0065" npl-type="s"><text>Cermak et al., Nucleic Acids Res 39, e82 (2011</text></nplcit>); <nplcit id="ncit0066" npl-type="s"><text>Wood et al., Science 333, 307 (2011</text></nplcit>); <nplcit id="ncit0067" npl-type="s"><text>Hockemeye et al. Nat Biotechnol 29, 731-734 (2011</text></nplcit>); <nplcit id="ncit0068" npl-type="s"><text>Tesson et al., Nat Biotechnol 29, 695-696 (2011</text></nplcit>); <nplcit id="ncit0069" npl-type="s"><text>Sander et al., Nat Biotechnol 29, 697-698 (2011</text></nplcit>); <nplcit id="ncit0070" npl-type="s"><text>Huang et al., Nat Biotechnol 29, 699-700 (2011</text></nplcit>); and <nplcit id="ncit0071" npl-type="s"><text>Zhang et al., Nat Biotechnol 29, 149-153 (2011</text></nplcit>.
0045Also suitable for use in the present methods are MegaTALs, which are a fusion of a meganuclease with a TAL effector; see, e.g., <nplcit id="ncit0072" npl-type="s"><text>Boissel et al., Nucl. Acids Res. 42(4):2591-2601 (2014</text></nplcit>); <nplcit id="ncit0073" npl-type="s"><text>Boissel and Scharenberg, Methods Mol Biol. 2015;1239:171-96</text></nplcit>.
Zinc Fingers
0046Zinc finger proteins are DNA-binding proteins that contain one or more zinc fingers, independently folded zinc-containing mini-domains, the structure of which is well known in the art and defined in, for example, <nplcit id="ncit0074" npl-type="s"><text>Miller et al., 1985, EMBO J., 4:1609</text></nplcit>; <nplcit id="ncit0075" npl-type="s"><text>Berg, 1988, Proc. Natl. Acad. Sci. USA, 85:99</text></nplcit>; <nplcit id="ncit0076" npl-type="s"><text>Lee et al., 1989, Science. 245:635</text></nplcit>; and <nplcit id="ncit0077" npl-type="s"><text>Klug, 1993, Gene, 135:83</text></nplcit>. Crystal structures of the zinc finger protein Zif268 and its variants bound to DNA show a semi-conserved pattern of interactions, in which typically three amino acids from the alpha-helix of the zinc finger contact three adjacent base pairs or a "subsite" in the DNA (<nplcit id="ncit0078" npl-type="s"><text>Pavletich et al., 1991, Science, 252:809</text></nplcit>; <nplcit id="ncit0079" npl-type="s"><text>Elrod-Erickson et al., 1998, Structure, 6:451</text></nplcit>). Thus, the crystal structure of Zif268 suggested that zinc finger DNA-binding domains might function in a modular manner with a one-to-one interaction between a zinc finger and a three-base-pair "subsite" in the DNA sequence. In naturally occurring zinc finger transcription factors, multiple zinc fingers are typically linked together in a tandem array to achieve sequence-specific recognition of a contiguous DNA sequence (<nplcit id="ncit0080" npl-type="s"><text>Klug, 1993, Gene 135:83</text></nplcit>).
0047Multiple studies have shown that it is possible to artificially engineer the DNA binding characteristics of individual zinc fingers by randomizing the amino acids at the alpha-helical positions involved in DNA binding and using selection methodologies such as phage display to identify desired variants capable of binding to DNA target sites of interest (<nplcit id="ncit0081" npl-type="s"><text>Rebar et al., 1994, Science, 263:671</text></nplcit>; <nplcit id="ncit0082" npl-type="s"><text>Choo et al., 1994 Proc. Natl. Acad. Sci. USA, 91:11163</text></nplcit>; <nplcit id="ncit0083" npl-type="s"><text>Jamieson et al., 1994, Biochemistry 33:5689</text></nplcit>; <nplcit id="ncit0084" npl-type="s"><text>Wu et al., 1995 Proc. Natl. Acad. Sci. USA, 92: 344</text></nplcit>). Such recombinant zinc finger proteins can be fused to functional domains, such as transcriptional activators, transcriptional repressors, methylation domains, and nucleases to regulate gene expression, alter DNA methylation, and introduce targeted alterations into genomes of model organisms, plants, and human cells (<nplcit id="ncit0085" npl-type="s"><text>Carroll, 2008, Gene Ther., 15:1463-68</text></nplcit>; <nplcit id="ncit0086" npl-type="s"><text>Cathomen, 2008, Mol. Ther., 16:1200-07</text></nplcit>; <nplcit id="ncit0087" npl-type="s"><text>Wu et al., 2007, Cell. Mol. Life Sci., 64:2933-44</text></nplcit>).
0048One existing method for engineering zinc finger arrays, known as "modular assembly," advocates the simple joining together of pre-selected zinc finger modules into arrays (<nplcit id="ncit0088" npl-type="s"><text>Segal et al., 2003, Biochemistry, 42:2137-48</text></nplcit>; <nplcit id="ncit0089" npl-type="s"><text>Beerli et al., 2002, Nat. Biotechnol., 20:135-141</text></nplcit>; <nplcit id="ncit0090" npl-type="s"><text>Mandell et al., 2006, Nucleic Acids Res., 34:W516-523</text></nplcit>; <nplcit id="ncit0091" npl-type="s"><text>Carroll et al., 2006, Nat. Protoc. 1:1329-41</text></nplcit>; <nplcit id="ncit0092" npl-type="s"><text>Liu et al., 2002, J. Biol. Chem., 277:3850-56</text></nplcit>; <nplcit id="ncit0093" npl-type="s"><text>Bae et al., 2003, Nat. Biotechnol., 21:275-280</text></nplcit>; <nplcit id="ncit0094" npl-type="s"><text>Wright et al., 2006, Nat. Protoc., 1:1637-52</text></nplcit>). Although straightforward enough to be practiced by any researcher, recent reports have demonstrated a high failure rate for this method, particularly in the context of zinc finger nucleases (<nplcit id="ncit0095" npl-type="s"><text>Ramirez et al., 2008, Nat. Methods, 5:374-375</text></nplcit>; <nplcit id="ncit0096" npl-type="s"><text>Kim et al., 2009, Genome Res. 19:1279-88</text></nplcit>), a limitation that typically necessitates the construction and cell-based testing of very large numbers of zinc finger proteins for any given target gene (<nplcit id="ncit0097" npl-type="s"><text>Kim et al., 2009, Genome Res. 19:1279-88</text></nplcit>).
0049Combinatorial selection-based methods that identify zinc finger arrays from randomized libraries have been shown to have higher success rates than modular assembly (<nplcit id="ncit0098" npl-type="s"><text>Maeder et al., 2008, Mol. Cell, 31:294-301</text></nplcit>; <nplcit id="ncit0099" npl-type="s"><text>Joung et al., 2010, Nat. Methods, 7:91-92</text></nplcit>; <nplcit id="ncit0100" npl-type="s"><text>Isalan et al., 2001, Nat. Biotechnol., 19:656-660</text></nplcit>). In preferred embodiments, the zinc finger arrays are described in, or are generated as described in, <patcit id="pcit0009" dnum="WO2011017293A"><text>WO 2011/017293</text></patcit> and <patcit id="pcit0010" dnum="WO2004099366A"><text>WO 2004/099366</text></patcit>. Additional suitable zinc finger DBDs are described in <patcit id="pcit0011" dnum="US6511808B"><text>U.S. Pat. Nos. 6,511,808</text></patcit>,<patcit id="pcit0012" dnum="US6013453A"><text> 6,013,453</text></patcit>, <patcit id="pcit0013" dnum="US6007988A"><text>6,007,988</text></patcit>, and <patcit id="pcit0014" dnum="US6503717B"><text>6,503,717</text></patcit> and <patcit id="pcit0015" dnum="US20020160940A" dnum-type="L"><text>U.S. patent application 2002/0160940</text></patcit>.
DNA
0050The methods described herein can be applied to any double stranded DNA, e.g., genomic DNA isolated from any cell, artificially created populations of DNAs, or any other DNA pools, as it is performed <i>in vitro.</i>
Sequencing
0051As used herein, "sequencing" includes any method of determining the sequence of a nucleic acid. Any method of sequencing can be used in the present methods, including chain terminator (Sanger) sequencing and dye terminator sequencing. In preferred embodiments, Next Generation Sequencing (NGS), a high-throughput sequencing technology that performs thousands or millions of sequencing reactions in parallel, is used. Although the different NGS platforms use varying assay chemistries, they all generate sequence data from a large number of sequencing reactions run simultaneously on a large number of templates. Typically, the sequence data is collected using a scanner, and then assembled and analyzed bioinformatically. Thus, the sequencing reactions are performed, read, assembled, and analyzed in parallel; see, e.g., <patcit id="pcit0016" dnum="US20140162897A"><text>US 20140162897</text></patcit>, as well as <nplcit id="ncit0101" npl-type="s"><text>Voelkerding et al., Clinical Chem., 55: 641-658, 2009</text></nplcit>; and <nplcit id="ncit0102" npl-type="s"><text>MacLean et al., Nature Rev. Microbiol., 7: 287-296 (2009</text></nplcit>). Some NGS methods require template amplification and some that do not. Amplification-requiring methods include pyrosequencing (see, e.g., <patcit id="pcit0017" dnum="US621089A"><text>U.S. Pat. Nos. 6,210,89</text></patcit> and <patcit id="pcit0018" dnum="US6258568B"><text>6,258,568</text></patcit>; commercialized by Roche); the Solexa/Illumina platform (see, e.g., <patcit id="pcit0019" dnum="US6833246B"><text>U.S. Pat. Nos. 6,833,246</text></patcit>, <patcit id="pcit0020" dnum="US7115400B"><text>7,115,400</text></patcit>, and <patcit id="pcit0021" dnum="US6969488B"><text>6,969,488</text></patcit>); and the Supported Oligonucleotide Ligation and Detection (SOLiD) platform (Applied Biosystems; see, e.g., <patcit id="pcit0022" dnum="US5912148A"><text>U.S. Pat. Nos. 5,912,148</text></patcit> and <patcit id="pcit0023" dnum="US6130073A"><text>6,130,073</text></patcit>). Methods that do not require amplification, e.g., single-molecule sequencing methods, include nanopore sequencing, HeliScope (<patcit id="pcit0024" dnum="US7169560B"><text>U.S. Pat. Nos. 7,169,560</text></patcit>; <patcit id="pcit0025" dnum="US7282337B"><text>7,282,337</text></patcit>; <patcit id="pcit0026" dnum="US7482120B"><text>7,482,120</text></patcit>; <patcit id="pcit0027" dnum="US7501245B"><text>7,501,245</text></patcit>; <patcit id="pcit0028" dnum="US6818395B"><text>6,818,395</text></patcit>; <patcit id="pcit0029" dnum="US6911345B"><text>6,911,345</text></patcit>; and <patcit id="pcit0030" dnum="US7501245B"><text>7,501,245</text></patcit>); real-time sequencing by synthesis (see, e.g., <patcit id="pcit0031" dnum="US7329492B"><text>U.S. Pat. No. 7,329,492</text></patcit>); single molecule real time (SMRT) DNA sequencing methods using zero-mode waveguides (ZMWs); and other methods, including those described in <patcit id="pcit0032" dnum="US7170050B"><text>U.S. Pat. Nos. 7,170,050</text></patcit>; <patcit id="pcit0033" dnum="US7302146B"><text>7,302,146</text></patcit>; <patcit id="pcit0034" dnum="US7313308B"><text>7,313,308</text></patcit>; and <patcit id="pcit0035" dnum="US7476503B"><text>7,476,503</text></patcit>). See, e.g., <patcit id="pcit0036" dnum="US20130274147A"><text>US 20130274147</text></patcit>; <patcit id="pcit0037" dnum="US20140038831A"><text>US20140038831</text></patcit>; <nplcit id="ncit0103" npl-type="s"><text>Metzker, Nat Rev Genet 11(1): 31-46 (2010</text></nplcit>).
0052Alternatively, hybridization-based sequence methods or other high-throughput methods can also be used, e.g., microarray analysis, NANOSTRING, ILLUMINA, or other sequencing platforms.
Kits
0053Also provided herein are kits for use in the methods described herein. The kits can include one or more of the following: 5' single-tailed hairpin adapters; 3' single-tailed hairpin adapters; reagents and/or enzymes for end repair and A tailing (e.g., T4 polymerase, Klenow fragment, T4 Polynucleotide Kinase (PNK), and/or Taq DNA Polymerase); exonuclease; uracil DNA glycosylase (UDG) and/or endonuclease VIII, a DNA glycosylase-lyase, e.g., the USER (Uracil-Specific Excision Reagent) Enzyme mixture (New England BioLabs); purified cas9 protein; guideRNA (e.g., control gRNA); gDNA template (e.g., control gDNA template); and instructions for use in a method described herein.
EXAMPLES
0054The invention is further described in the following examples, which do not limit the scope of the invention described in the claims.
Materials
0055The following materials were used in Examples 1-2. <tables id="tabl0001" num="0001"><table frame="topbot"><tgroup cols="3" colsep="0"><colspec colnum="1" colname="col1" colwidth="69mm" /><colspec colnum="2" colname="col2" colwidth="31mm" /><colspec colnum="3" colname="col3" colwidth="27mm" /><thead><row><entry align="center"><b>Materials</b></entry><entry align="center"><b>Vendor</b></entry><entry align="center"><b>Model Number</b></entry></row></thead><tbody><row rowsep="0"><entry align="center" valign="bottom">HTP Library Preparation Kit</entry><entry align="center" valign="bottom">Kapa Biosystems</entry><entry align="center" valign="bottom">KK8235</entry></row><row><entry align="center" valign="bottom">Hifi HotStart ReadyMix, 100 x 25 µL reactions</entry><entry align="center" valign="bottom">Kapa Biosystems</entry><entry align="center" valign="bottom">KK2602</entry></row><row rowsep="0"><entry align="center" valign="bottom">Lambda exonuclease (5U/ul)</entry><entry align="center" valign="bottom">NEB</entry><entry align="center" valign="bottom">M0262L</entry></row><row rowsep="0"><entry align="center" valign="bottom">E. Coli Exonuclease I (20U/ul)</entry><entry align="center" valign="bottom">NEB</entry><entry align="center" valign="bottom">M0293S</entry></row><row rowsep="0"><entry align="center" valign="bottom">USER enzyme (1000U/ul)</entry><entry align="center" valign="bottom">NEB</entry><entry align="center" valign="bottom">M5505L</entry></row><row><entry align="center" valign="bottom">Cas9 enzyme (1000 nM)</entry><entry align="center" valign="bottom">NEB</entry><entry align="center" valign="bottom">M0386L</entry></row><row><entry align="center" valign="bottom">Ampure XP 60 ml</entry><entry align="center" valign="bottom">Agencourt</entry><entry align="center" valign="bottom">A63881</entry></row></tbody></tgroup></table></tables>
0056The following Primers were used in Examples 1-2. <tables id="tabl0002" num="0002"><table frame="topbot"><tgroup cols="3" colsep="0"><colspec colnum="1" colname="col1" colwidth="58mm" /><colspec colnum="2" colname="col2" colwidth="95mm" /><colspec colnum="3" colname="col3" colwidth="13mm" /><thead><row><entry namest="col1" nameend="col3" align="center" valign="top"><b>TABLE A</b></entry></row><row><entry valign="top"><b>PrimerName</b></entry><entry valign="top"><b>Sequence</b></entry><entry valign="top"><b>SEQ ID NO:</b></entry></row></thead><tbody><row><entry>oSQT1270 5'-Truseq-loop-adapter D501L</entry><entry><img file="EP3347467B1_D0001.tif" /></entry><entry>1</entry></row><row><entry>oSQT1302 5'-Truseq-loop-adapter D502L</entry><entry><img file="EP3347467B1_D0002.tif" /></entry><entry>2</entry></row><row><entry>oSQT1303 5'-Truseq-loop-adapter D503L</entry><entry><img file="EP3347467B1_D0003.tif" /></entry><entry>3</entry></row><row><entry>oSQT1304 5'-Truseq-loop-adapter D504L</entry><entry><img file="EP3347467B1_D0004.tif" /></entry><entry>4</entry></row><row><entry>oSQT1305 5'-Truseq-loop-adapter D505L</entry><entry><img file="EP3347467B1_D0005.tif" /></entry><entry>5</entry></row><row><entry>oSQT1315 5'-Truseq-loop-adapter D506L</entry><entry><img file="EP3347467B1_D0006.tif" /></entry><entry>6</entry></row><row><entry>oSQT1316 5'-Truseq-loop-adapter D507L</entry><entry><img file="EP3347467B1_D0007.tif" /></entry><entry>7</entry></row><row><entry>oSQT1317 5'-Truseq-loop-adapter D508L</entry><entry><img file="EP3347467B1_D0008.tif" /></entry><entry>8</entry></row><row><entry>oSQT1271 3'-Truseq-loop-adapter D701L</entry><entry><img file="EP3347467B1_D0009.tif" /></entry><entry>9</entry></row><row><entry>oSQT1306 3'-Truseq-loop-adapter D702L</entry><entry><img file="EP3347467B1_D0010.tif" /></entry><entry>10</entry></row><row><entry>oSQT1307 3'-Truseq-loop-adapter D703L</entry><entry><img file="EP3347467B1_D0011.tif" /></entry><entry>11</entry></row><row><entry>oSQT1308 3'-Truseq-loop-adapter D704L</entry><entry><img file="EP3347467B1_D0012.tif" /></entry><entry>12</entry></row><row><entry>oSQT1309 3'-Truseq-loop-adapter D705L</entry><entry><img file="EP3347467B1_D0013.tif" /></entry><entry>13</entry></row><row><entry>oSQT1318 3'-Truseq-loop-adapter D706L</entry><entry><img file="EP3347467B1_D0014.tif" /></entry><entry>14</entry></row><row><entry>oSQT1319 3'-Truseq-loop-adapter D707L</entry><entry><img file="EP3347467B1_D0015.tif" /></entry><entry>15</entry></row><row><entry>oSQT1320 3'-Truseq-loop-adapter D708L</entry><entry><img file="EP3347467B1_D0016.tif" /></entry><entry>16</entry></row><row><entry>oSQT1274 Truseq F1</entry><entry>AATGATACGGCGACCACCGAG</entry><entry>17</entry></row><row><entry>oSQT1275 Truseq R1</entry><entry>CAAGCAGAAGACGGCATACGAGAT</entry><entry>18</entry></row></tbody></tgroup></table></tables>
0057The sequences in CAPS are dual-index barcodes (for demultiplexing samples) in the Illumina Truseq library preparation system. The * represents a phosphorothioate linkage.
Example 1. Optimization of Exemplary FIND-Seq Methodology
0058Described herein is the development of an <i>in vitro</i> method that enables comprehensive determination of DSBs induced by nucleases on any genomic DNA of interest. This method enables enrichment of nuclease-induced DSBs over random DSBs induced by shearing of genomic DNA. An overview of how the method works can be found in <figref idref="f0001"><b>Figures 1A</b></figref> and <figref idref="f0002"><b>1B</b></figref><b>.</b> In brief, genomic DNA from a cell type or organism of interest is randomly sheared to a defined average length. In the present experiments, an average length of 500 bps works well for subsequent next-generation sequencing steps. The random broken ends of this genomic DNA are end repaired and then A-tailed. This then enables the ligation of a hairpin adapter bearing a 5' end next-generation sequencing primer site that also contains a single deoxyuridine nucleotide at a specific position <b>(</b><figref idref="f0002"><b>Figure 1B</b></figref><b>);</b> this adapter is also referred to herein as the "5' single-tailed hairpin adapter". Following ligation of this adapter, the sample is then treated with one or more exonucleases (e.g., bacteriophage lambda exonuclease, E. coli ExoI, PlasmidSafe ATP-dependent exonuclease), which degrades any DNA molecules that have not had the 5' single-tailed hairpin adapter ligation to both of their ends. The collection of molecules is then treated with Cas9 nuclease complexed with a specific guide RNA (gRNA). The resulting nuclease-induced blunt ends are A-tailed and then a second hairpin adapter bearing a 3' end next-generation sequencing primer site and that also contains a single deoxyuridine nucleotide at a specific position is ligated to these ends; this adapter is also referred to herein as the "3' single-tailed hairpin adapter". As a result of these treatments, the only DNA fragments that should have a 5' and a 3' single-tailed hairpin adapter ligated to their ends are those that were cleaved by the nuclease. The adapters can be added in any order. Following treatment by the USER enzyme mixture, which nicks DNA wherever a deoxyuridine is present, only those fragments bearing the two types of adapters can be sequenced.
0059Following paired end next-generation sequencing, the resulting reads can be mapped back to the genome. Sites of nuclease cleavage will have multiple bi-direction reads originating at the site of the nuclease-induced DSB, typically within a nucleotide of the cut site <b>(</b><figref idref="f0003"><b>Figures 2A</b></figref><b>and</b><figref idref="f0004"><b>2B</b></figref><b>).</b> As nuclease-induced cleavage produces 'uniform' ends, in contrast to the staggered ends that results from random physical shearing of DNA, these sites can be simply bioinformatically identified by their signature uniform read alignments.
0060Using an initial non-optimized version of this approach, off-target sites induced by four different gRNAs and Cas9 nuclease were successfully identified. The target sites are listed in Table 1. <tables id="tabl0003" num="0003"><table frame="none"><title><b>Table 1. List of sgRNA target sites tested with <i>in vitro</i> cleavage selection assay.</b></title><tgroup cols="4" colsep="0" rowsep="0"><colspec colnum="1" colname="col1" colwidth="29mm" /><colspec colnum="2" colname="col2" colwidth="19mm" /><colspec colnum="3" colname="col3" colwidth="64mm" /><colspec colnum="4" colname="col4" colwidth="23mm" /><thead><row><entry align="center" valign="top"><b>Target site name</b></entry><entry align="center" valign="top"><b>Cells</b></entry><entry align="center" valign="top"><b>Target Site (5' → 3')</b></entry><entry align="center" valign="top"><b>SEQ ID NO:</b></entry></row></thead><tbody><row><entry align="center">VEGFA <b>site1</b></entry><entry align="center">U20S</entry><entry align="center">GGGTGGGGGGAGTTTGCTCCNGG</entry><entry align="center">19</entry></row><row><entry align="center">VEGFA site2</entry><entry align="center">U20S</entry><entry align="center">GACCCCCTCCACCCCGCCTCNGG</entry><entry align="center">20</entry></row><row><entry align="center">VEGFA site3</entry><entry align="center">U20S</entry><entry align="center">GGTGAGTGAGTGTGTGCGTGNGG</entry><entry align="center">21</entry></row><row><entry align="center">EMX1</entry><entry align="center">U20S</entry><entry align="center">GAGTCCGAGCAGAAGAAGAANGG</entry><entry align="center">22</entry></row></tbody></tgroup></table></tables>
0061The results of table 2 were encouraging for this assay, as the majority of GUIDE-seq detected sites were also detected by this method. In the first trial, 58-77% of sites overlapped between methods using 2.5-3M reads. ∼20,000X enrichment was estimated for cleaved sites. <tables id="tabl0004" num="0004"><table frame="topbot"><title><b>Table 2. List of sgRNA target sites tested with <i>in vitro</i> cleavage selection assay.</b></title><tgroup cols="5" colsep="0"><colspec colnum="1" colname="col1" colwidth="24mm" /><colspec colnum="2" colname="col2" colwidth="39mm" /><colspec colnum="3" colname="col3" colwidth="17mm" /><colspec colnum="4" colname="col4" colwidth="19mm" /><colspec colnum="5" colname="col5" colwidth="68mm" /><thead><row><entry align="center"><b>site</b></entry><entry align="center"><b>cell-based GUIDE-seq total</b></entry><entry align="center"><b>in vitro total</b></entry><entry align="center"><b>both</b></entry><entry align="center"><b>percentage</b> of <b>cell-based GUIDE-seq sites detected</b></entry></row></thead><tbody><row rowsep="0"><entry align="center" valign="bottom"><i>EMX1</i></entry><entry align="center" valign="bottom">16</entry><entry align="center" valign="bottom">19</entry><entry align="center" valign="bottom">10</entry><entry align="center" valign="bottom">63%</entry></row><row rowsep="0"><entry align="center" valign="bottom"><i>VEGFA</i> site 1</entry><entry align="center" valign="bottom">22</entry><entry align="center" valign="bottom">41</entry><entry align="center" valign="bottom">17</entry><entry align="center" valign="bottom">77%</entry></row><row rowsep="0"><entry align="center" valign="bottom"><i>VEGFA</i> site 2</entry><entry align="center" valign="bottom">152</entry><entry align="center" valign="bottom">194</entry><entry align="center" valign="bottom">98</entry><entry align="center" valign="bottom">64%</entry></row><row><entry align="center" valign="bottom"><i>VEGFA</i> site 3</entry><entry align="center" valign="bottom">60</entry><entry align="center" valign="bottom">129</entry><entry align="center" valign="bottom">35</entry><entry align="center" valign="bottom">58%</entry></row></tbody></tgroup></table></tables>
0062A critical parameter for optimization of the present method was the efficiency of degradation of linear DNA fragments following the ligation of the first 5' single-tailed hairpin adapter. Indeed, without wishing to be bound by theory, most of the "background" reads mapped throughout the genome may be due to incomplete digestion of unligated DNA at this step, thereby leaving ends to which the second 3' single tailed hairpin adapter can be ligated.
0063Therefore, the methods were further optimized by testing various kinds and numbers of exonucleases and lengths of incubation to find an optimal treatment method. As seen in <figref idref="f0005 f0006"><b>Figures 3A-3D</b></figref><b>,</b> three treatments consistently yielded the highest number of reads and sites and also found 100% of the off-target sites identified by matched GUIDE-seq experiments with the same gRNAs: 1 hour PlasmidSafe treatment, two serial 1 hour PlasmidSafe treatments, and 1 hour of Lambda exonuclease treatment followed by 1 hour of PlasmidSafe treatment.
0064Using an optimized exonuclease treatment, the methods were again performed using CRISPR-Cas9 nuclease and a gRNA targeted to VEGFA site 1. This experiment yielded a larger number of off-target sites. Combining the various experiments performed with the VEGFA site 1 gRNA yields a larger number of off-target sites. GUIDE-seq had previously been performed using this same VEGFA site 1 gRNA and a range of genome-wide off-target sites identified. Importantly, the present in vitro method performed with the optimized exonuclease treatment found ALL of the off-target sites previously identified by GUIDE-seq for the gRNA tested and a very large number of additional, previously unknown off-target sites as well <b>(</b><figref idref="f0007"><b>Figure 4</b></figref><b>)</b>. This result demonstrates that these new methods are as sensitive as GUIDE-seq but offers the ability to identify sites that are not found by that earlier method. Reasons for this might include greater sensitivity of this in vitro method, negative biological selection against certain off-target mutations when practicing GUIDE-seq in cells, and/or negative effects of chromatin, DNA methylation, and gene expression on the activity of nucleases in cells.
Example 2. Exemplary FIND-Seq protocol
0065Selection of Cleaved Amplicons for Next-generation Sequencing (SCANS) <i>in vitro assay for finding cleavage sites of CRISPR</i>/<i>Cas9 nuclease from complex genomic DNA mixtures</i><tables id="tabl0005" num="0005"><table frame="top"><tgroup cols="3" colsep="0" rowsep="0"><colspec colnum="1" colname="col1" colwidth="84mm" /><colspec colnum="2" colname="col2" colwidth="18mm" /><colspec colnum="3" colname="col3" colwidth="14mm" /><thead><row><entry valign="top">1. End Repair (Red)</entry><entry align="center" valign="top" /><entry align="center" valign="top" /></row></thead><tbody><row><entry>Component</entry><entry align="center">1 rxn (ul)</entry><entry align="center">12</entry></row><row><entry>Water</entry><entry align="center">8</entry><entry align="center">115</entry></row><row><entry>10X Kapa End Repair Buffer</entry><entry align="center">7</entry><entry align="center">101</entry></row><row><entry>Kapa End Repair Enzyme Mix</entry><entry align="center">5</entry><entry align="center">72</entry></row><row><entry>Total master mix volume</entry><entry align="center">20</entry><entry align="center">288</entry></row><row><entry>Input DNA (0.1 - 5 ug)</entry><entry align="center">50</entry><entry align="center" /></row><row><entry>Total reaction volume</entry><entry align="center">70</entry><entry align="center" /></row><row><entry>Incubate 30 minutes at 20C.</entry><entry align="center" /><entry align="center" /></row><row><entry>1.7X SPRI cleanup using 120 ul Ampure XP beads.</entry><entry align="center" /><entry align="center" /></row><row><entry>2. A-tailing (Blue)</entry><entry align="center" /><entry align="center" /></row><row><entry>Component</entry><entry align="center">1 rxn (ul)</entry><entry align="center">12</entry></row><row><entry>10X Kapa A-Tailing Buffer</entry><entry align="center">5</entry><entry align="center">72</entry></row><row><entry>Kapa A-Tailing Enzyme</entry><entry align="center">3</entry><entry align="center">43</entry></row><row><entry>Total master mix volume</entry><entry align="center">8</entry><entry align="center">115</entry></row><row><entry>TE (0.1 mM EDTA)</entry><entry align="center">42</entry><entry align="center" /></row><row><entry>End repaired DNA with beads</entry><entry align="center">0</entry><entry align="center" /></row><row><entry>Total reaction volume</entry><entry align="center">50</entry><entry align="center" /></row><row><entry>Incubate for 30 min at 30C.</entry><entry align="center" /><entry align="center" /></row><row><entry>Cleanup by adding 90 ul of SPRI solution.</entry><entry align="center" /><entry align="center" /></row><row><entry>3. 5' Adapter Ligation (Yellow)</entry><entry align="center" /><entry align="center" /></row><row><entry>Component</entry><entry align="center">1 rxn (ul)</entry><entry align="center">12</entry></row><row><entry>5X Kapa Ligation Buffer</entry><entry align="center">10</entry><entry align="center">144</entry></row><row><entry>Kapa T4 DNA Ligase</entry><entry align="center">5</entry><entry align="center">72</entry></row><row><entry>Total master mix volume</entry><entry align="center">15</entry><entry align="center">216</entry></row><row><entry>A-tailed DNA with beads</entry><entry align="center">0</entry><entry align="center" /></row><row><entry>TE (0.1 mM EDTA)</entry><entry align="center">30</entry><entry align="center" /></row><row><entry>5' Truseq Loop Adapter D50X (40 uM)</entry><entry align="center">5</entry><entry align="center" /></row><row><entry>Total reaction volume</entry><entry align="center">50</entry><entry align="center" /></row><row><entry>Incubate for 1 hour at 20C.</entry><entry align="center" /><entry align="center" /></row><row><entry>Cleanup by adding 50 ul of SPRI solution.</entry><entry align="center" /><entry align="center" /></row><row><entry>Proceed with 1 ug of 5' Adapter-ligated DNA into next step.</entry><entry align="center" /><entry align="center" /></row><row><entry>4. Exonuclease treatment</entry><entry align="center" /><entry align="center" /></row><row><entry>Component</entry><entry align="center">1 rxn (ul)</entry><entry align="center">12</entry></row><row><entry>10X ExoI buffer</entry><entry align="center">5</entry><entry align="center">72</entry></row><row><entry>Lambda exonuclease (5 U/ul)</entry><entry align="center">4</entry><entry align="center">58</entry></row><row><entry>E. Coli Exonuclease I (20 U/ul)</entry><entry align="center">1</entry><entry align="center">14</entry></row><row><entry>Total master mix volume</entry><entry align="center">10</entry><entry align="center">144</entry></row><row><entry>Adapter-ligated DNA with beads</entry><entry align="center">0</entry><entry align="center" /></row><row><entry>TE (0.1 mM EDTA)</entry><entry align="center">40</entry><entry align="center" /></row><row><entry>Total reaction volume</entry><entry align="center">50</entry><entry align="center" /></row><row><entry>Incubate for 1 hour at 37C.</entry><entry align="center" /><entry align="center" /></row><row><entry>Cleanup by adding 50 ul of 1X Ampure XP beads.</entry><entry align="center" /><entry align="center" /></row><row><entry>5. Cleavage by Cas9</entry><entry align="center" /><entry align="center" /></row><row><entry>Component</entry><entry align="center">1 rxn (ul)</entry><entry align="center">5</entry></row><row><entry>Water</entry><entry align="center">63</entry><entry align="center">378</entry></row><row><entry>10X Cas9 buffer</entry><entry>10</entry><entry>60</entry></row><row><entry>Cas9 (1 uM -> 900 nM final)</entry><entry>9</entry><entry>54</entry></row><row><entry>sgRNA (300 nM, ∼100 ng/ul)</entry><entry>3</entry><entry>18</entry></row><row><entry>Total reaction volume</entry><entry>85</entry><entry>510</entry></row><row><entry>Incubate at room temperature for 10 minutes.</entry><entry /><entry /></row><row><entry>DNA (∼400 bp, 250 ng)</entry><entry>15</entry><entry /></row><row><entry>Incubate for 1 hour at 37C.</entry><entry /><entry /></row><row><entry>Purify with 1X SPRI bead cleanup (100 ul).</entry><entry /><entry /></row><row><entry>6. A-tailing</entry><entry /><entry /></row><row><entry>Component</entry><entry>1 rxn (ul)</entry><entry>5</entry></row><row><entry>10X Kapa A-Tailing Buffer</entry><entry>5</entry><entry>30</entry></row><row><entry>Kapa A-Tailing Enzyme</entry><entry>3</entry><entry>18</entry></row><row><entry>Total master mix volume</entry><entry>8</entry><entry>48</entry></row><row><entry>TE (0.1 mM EDTA)</entry><entry>42</entry><entry /></row><row><entry>End repaired DNA with beads</entry><entry>0</entry><entry /></row><row><entry>Total reaction volume</entry><entry>50</entry><entry /></row><row><entry>Incubate for 30 min at 30C.</entry><entry /><entry /></row><row><entry>Cleanup by adding 90 ul of SPRI solution.</entry><entry /><entry /></row><row><entry>7. 3' Adapter Ligation</entry><entry /><entry /></row><row><entry>Component</entry><entry>1 rxn (ul)</entry><entry>5</entry></row><row><entry>5X Kapa Ligation Buffer</entry><entry>10</entry><entry>60</entry></row><row><entry>Kapa T4 DNA Ligase</entry><entry>5</entry><entry>30</entry></row><row><entry>Total master mix volume</entry><entry>15</entry><entry>90</entry></row><row><entry>A-tailed DNA with beads</entry><entry>0</entry><entry /></row><row><entry>TE (0.1 mM EDTA)</entry><entry>30</entry><entry /></row><row><entry>3' Truseq Loop Adapter D70X (40 uM)</entry><entry>5</entry><entry /></row><row><entry>Total reaction volume</entry><entry>50</entry><entry /></row><row><entry>Incubate for 30 minutes at 20C.</entry><entry /><entry /></row><row><entry>Cleanup by adding 50 ul of SPRI solution.</entry><entry /><entry /></row><row><entry>8. USER enzyme treatment</entry><entry /><entry /></row><row><entry>Add 3 ul of USER enzyme and treat for 30 minutes at 37C.</entry><entry /><entry /></row><row><entry>(Treatment is in TE 10 mM Tris, 0.1 mM</entry><entry /><entry /></row><row><entry>EDTA.)</entry><entry /><entry /></row><row><entry>Purify with 1X SPRI solution (50 ul).</entry><entry /><entry /></row><row><entry>9. PCR amplification</entry><entry /><entry /></row><row><entry namest="col1" nameend="col3" align="left">Amplify with Kapa Hifi manufacturer's protocol and primers oSQT1274/1275</entry></row></tbody></tgroup></table></tables>
OTHER EMBODIMENTS
0066It is to be understood that while the invention has been described in conjunction with the detailed description thereof, the foregoing description is intended to illustrate and not limit the scope of the invention, which is defined by the scope of the appended claims. Other aspects, advantages, and modifications are within the scope of the following claims.
Contents8
24 sheets
Sheet 1 Sheet 2 Sheet 3 Sheet 4 Sheet 5 Sheet 6 Sheet 7 Sheet 8 Sheet 9 Sheet 10 Sheet 11 Sheet 12 Sheet 13 Sheet 14 Sheet 15 Sheet 16 Sheet 17 Sheet 18 Sheet 19 Sheet 20 Sheet 21 Sheet 22 Sheet 23 Sheet 24
Every citation, both ways
| Document | Relation | Office |
|---|---|---|
| WO2014071070A1 | Cites | World Intellectual Property Organization (WIPO) |
| US2006292611A1 | Cites | United States of America |
| US2013309668A1 | Cites | United States of America |
| US2014295557A1 | Cites | United States of America |
| SHENGDAR Q TSAI ET AL: "GUIDE-seq enables genome-wide profiling of off-target cleavage by CRISPR-Cas nucleases", NATURE BIOTECHNOLOGY, vol. 33, no. 2, 16 December 2014 (2014-12-16), pages 187-197, XP055246459, ISSN: 1087-0156, DOI: 10.1038/nbt.3117 | Non-patent | – |
16 members in 9 offices
Priority claims7
| Document | Office | Kind | Date |
|---|---|---|---|
| 201562217690 | United States of America | P | |
| 201562217690P | United States of America | – | |
| 2016051097 | United States of America | W | |
| 201562217690P | – | – | – |
| US201562217690P | – | – | – |
| US2016051097 | – | – | – |
| WO2016US51097 | – | – | – |
Members16
| Document | Office | Kind | |
|---|---|---|---|
| CA3000816A1 | Canada | A1 | |
| US2017073747A1 | United States of America | A1 | |
| WO2017044843A1 | World Intellectual Property Organization (WIPO) | A1 | |
| AU2016319110A1 | Australia | A1 | |
| KR20180043369A | Republic of Korea | A | |
| IL257955A | Israel | A | |
| IL257955D0 | Israel | D0 | |
| US9988674B2 | United States of America | B2 | |
| EP3347467A1 | European Patent Office (EPO) | A1 | |
| CN108350453A | China | A | |
| US2018265920A1 | United States of America | A1 | |
| JP2018530536A | Japan | A | |
| EP3347467A4 | European Patent Office (EPO) | A4 | |
| US11028429B2 | United States of America | B2 | |
| EP3347467B1This record | European Patent Office (EPO) | B1 | |
| AU2016319110B2 | Australia | B2 |
78 legal events, as 9 offices reported them to INPADOC
Over the term
Point at a mark for the eventEvents
| Event | Code | Office | |
|---|---|---|---|
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Gb: european patent ceased through non-payment of renewal feeCeasedGBPC | GBPC | EP | |
| Application deemed withdrawn, or ip right lapsed, due to non-payment of renewal feeWithdrawnR119 | R119 | DE | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| No opposition filedOpposition26N | 26N | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed because of non-payment of the annual feeLapsedMM | MM | BE | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Patent ceasedCeasedPL | PL | CH | |
| No opposition filed within time limitOppositionORIGINAL CODE: 0009261PLBE | PLBE | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: NO OPPOSITION FILED WITHIN TIME LIMITSTAA | STAA | EP | |
| No opposition filed against granted patent, or epo opposition proceedings concluded without decisionGrantedR097 | R097 | DE | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Patent invalid in the netherlands as no translation has been filedMP | MP | NL | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Deletion acc. to par. 5 (withdrawal of the translation of the ep patent)MK05 | MK05 | AT | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Invalidation of extension of european patentsMG9D | MG9D | LT | |
| European patents granted designating irelandGrantedFG4D | FG4D | IE | |
| Dpma publication of mentioned ep patent grantGrantedR096 | R096 | DE | |
| Reference to at number (ep patent validated in austria)REF | REF | AT | |
| European patent takes effect as a national patent in ch/liEP | EP | CH | |
| Designated contracting statesAK | AK | EP | |
| European patent grantedGrantedFG4D | FG4D | GB | |
| (expected) grantORIGINAL CODE: 0009210GRAA | GRAA | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: THE PATENT HAS BEEN GRANTEDSTAA | STAA | EP | |
| Grant fee paidORIGINAL CODE: EPIDOSNIGR3GRAS | GRAS | EP | |
| Intention to grant announcedINTG | INTG | EP | |
| Information provided on ipc code assigned before grantRIC1 | RIC1 | EP | |
| Information provided on ipc code assigned before grantRIC1 | RIC1 | EP | |
| Despatch of communication of intention to grant a patentORIGINAL CODE: EPIDOSNIGR1GRAP | GRAP | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: GRANT OF PATENT IS INTENDEDSTAA | STAA | EP | |
| First examination report despatched17Q | 17Q | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: EXAMINATION IS IN PROGRESSSTAA | STAA | EP | |
| Supplementary search report drawn up and despatchedA4 | A4 | EP | |
| Information provided on ipc code assigned before grantRIC1 | RIC1 | EP | |
| Information provided on ipc code assigned before grantRIC1 | RIC1 | EP | |
| Information provided on ipc code assigned before grantRIC1 | RIC1 | EP | |
| Request for validation of the european patent (deleted)DAV | DAV | EP | |
| Request for extension of the european patent (deleted)DAX | DAX | EP | |
| Request for examination filed17P | 17P | EP | |
| Designated contracting statesAK | AK | EP | |
| Request for extension of the european patentAX | AX | EP | |
| Public reference made under article 153(3) epc to a published international application that has entered the european phaseORIGINAL CODE: 0009012PUAI | PUAI | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: REQUEST FOR EXAMINATION WAS MADESTAA | STAA | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: THE INTERNATIONAL PUBLICATION HAS BEEN MADESTAA | STAA | EP |
Numbers
- Publication
- 3347467
- Publication, DOCDB
- 3347467
- Publication, EPODOC
- EP3347467
- Application
- 168451839
- Application, DOCDB
- 16845183
- Application, EPODOC
- EP20160845183
Titles3
- German
- VOLLSTÄNDIGE ABFRAGE DER NUKLEASE DSBS UND SEQUENZIERUNG (FIND-SEQ)
- English
- FULL INTERROGATION OF NUCLEASE DSBS AND SEQUENCING (FIND-SEQ)
- French
- INTERROGATION COMPLÈTE DE DSB NUCLÉASIQUES ET SÉQUENÇAGE (FIND-SEQ)
Classification
- CPC, 10
- C12N15/1093
- C12Q1/6855
- C12Q2521/301
- C12Q2521/319
- C12Q2521/501
- C12Q2525/191
- C12Q2525/301
- C12Q2535/122
- C40B50/06
- C40B40/06
- IPC, 2
- C12N15 10
- C12Q1 6855
Designated states1
- Contracting states, 1
- Türkiye
