DETAILED ACTION
Notice of Pre-AIA or AIA Status
The present application, filed on or after March 16, 2013, is being examined under the first inventor to file provisions of the AIA .
Claim Rejections - 35 USC § 102
The following is a quotation of the appropriate paragraphs of 35 U.S.C. 102 that form the basis for the rejections under this section made in this Office action:
A person shall be entitled to a patent unless –
(a)(1) the claimed invention was patented, described in a printed publication, or in public use, on sale, or otherwise available to the public before the effective filing date of the claimed invention.
Claim(s) 1-9 is/are rejected under 35 U.S.C. 102(a)(1) as being anticipated by Arik et al. (PGPUB 2018/0247636), hereinafter referenced as Arik.
Regarding claim 1, Arik discloses a method of speech synthesis for an audio signal, the method comprising:
receiving audio parameters of the audio signal and a representation of the audio signal (fig. 1, element 105, 135, 140 with spectral or excitation parameters; p. 0031);
inputting the audio parameters of the audio signal and the representation of the audio signal into a generative model (generative model), wherein the generative model is trained to generate a synthesized representation of the audio signal based on the audio parameters and the representation of the audio signal (fig. 1 with p. 0058-0061, 0092-0100 with wavenet audio synthesis); and
outputting, by the generative model, the synthesized representation of the audio signal (final synthesized utterances; p. 0041-0042).
Regarding claims 2 and 8, Arik discloses a method wherein the audio signal comprises 16 kHz audio (p. 0095).
Regarding claims 3 and 7, Arik discloses a method wherein the generative model is configured to operate autoregressively (p. 0025, 0059, 0100, 0166).
Regarding claims 4 and 9, Arik discloses a method wherein the generative model comprises a generative adversarial network (p. 0128).
Regarding claim 5, Arik discloses a method further comprising quantizing at least one of the representation of the audio signal or the audio parameters of the audio signal (quantize; p. 0128, 0135).
Regarding claim 6, it interpreted and rejected for similar reasons as set forth above. In addition, Arik discloses a system for generating a synthesized representation of an audio signal, the system comprising:
a generative model (p. 0092, 0128 and 0163) comprising:
a layer configured to use a tanh activation (p. 0173); and
a convolutional layer (p. 0048-0060).
Conclusion
Any inquiry concerning this communication or earlier communications from the examiner should be directed to JAKIEDA R JACKSON whose telephone number is (571)272-7619. The examiner can normally be reached Mon - Fri 6:30a-2:30p.
Examiner interviews are available via telephone, in-person, and video conferencing using a USPTO supplied web-based collaboration tool. To schedule an interview, applicant is encouraged to use the USPTO Automated Interview Request (AIR) at http://www.uspto.gov/interviewpractice.
If attempts to reach the examiner by telephone are unsuccessful, the examiner’s supervisor, Daniel Washburn can be reached at 571.272.5551. The fax phone number for the organization where this application or proceeding is assigned is 571-273-8300.
Information regarding the status of published or unpublished applications may be obtained from Patent Center. Unpublished application information in Patent Center is available to registered users. To file and manage patent submissions in Patent Center, visit: https://patentcenter.uspto.gov. Visit https://www.uspto.gov/patents/apply/patent-center for more information about Patent Center and https://www.uspto.gov/patents/docx for information about filing in DOCX format. For additional questions, contact the Electronic Business Center (EBC) at 866-217-9197 (toll-free). If you would like assistance from a USPTO Customer Service Representative, call 800-786-9199 (IN USA OR CANADA) or 571-272-1000.
/JAKIEDA R JACKSON/Primary Examiner, Art Unit 2657