MPEP § 2412.05(a) — Use of Sequentially Numbered Sequence Identifiers in the “Sequence Listing XML”
Ninth Edition, Revision 01.2024 · last revised R-01.2024
[Editor Note: This section is applicable to all applications with a filing date, or, for national phase applications, an international filing date, on or after July 1, 2022, having disclosure of one or more nucleotide and/or amino acid sequences as defined in 37 CFR 1.831(b). Formatting representations of XML (eXtensible Markup Language) elements in this section appear different than shown in Standard ST.26, which may be accessed at: www.wipo.int /export/sites/www/standards/en/pdf/03-26-01.pdf.]
37 CFR 1.832 Representation of nucleotide and/or amino acid sequence data in the “Sequence Listing XML” part of a patent application filed on or after July 1, 2022.
- (a) Each disclosed nucleotide or amino acid sequence that meets the requirements of § 1.831(b) must appear separately in the “Sequence Listing XML.” Each sequence set forth in the “Sequence Listing XML” must be assigned a separate sequence identifier. The sequence identifiers must begin with 1 and increase sequentially by integers as defined in paragraph 10 of WIPO Standard ST.26 (incorporated by reference, see § 1.839).
-
*****
In accordance with 37 CFR 1.832(a), the sequence identifiers in the “Sequence Listing XML” must begin with 1 and increase sequentially by integers. The requirement for sequence identifiers, at a minimum, requires that each sequence be assigned a different number for purposes of identification. However, where practical and for ease of reference, sequences should be presented in the “Sequence Listing XML” in numerical order and in the order in which they are discussed in the application.
Each nucleotide and/or amino acid sequence that meets the definition in 37 CFR 1.831(b) and is enumerated by its residues must be assigned a separate sequence identifier, including a sequence which is identical to a region of a longer sequence. See MPEP § 2412.02 for further description of a “sequence”.
Where no sequence is present for a sequence identifier, i.e. an intentionally skipped sequence, “000” must be used in place of a sequence in a “Sequence Listing XML”. The total number of sequences indicated in the “Sequence Listing XML” must equal the total number of sequence identifiers, whether followed by a sequence or by “000”.
WIPO Standard ST.26 paragraphs 58 and 59 require that an intentionally skipped sequences in a “Sequence Listing XML” must be represented as follows:
- (a) the value of the element SequenceData and its attribute sequenceIDNumber, is the sequence identifier of the skipped sequence;
- (b) no value is provided for the elements INSDSeq _length, INSDSeq _moltype, and INSDSeq _division;
- (c) the element INSDSeq _feature-table is not included; and
- (d) the value of the element INSDSeq _sequence is the string “000”.
Cited authority
- 37 CFR 1.831 Requirements for patent applications filed on or after July 1, 2022, having nucleotide and/or amino acid sequence disclosures
- 37 CFR 1.839 Incorporation by reference
- 37 CFR 1.832 Representation of nucleotide and/or amino acid sequence data in the “Sequence Listing XML” part of a patent application filed on or after July 1, 2022
- 2412.02 Definition of “Sequence Listing XML”