G10L21/0216

CAMERA-VIEW ACOUSTIC FENCE
20230053202 · 2023-02-16 ·

Determining the angle of sound relative to the centerline of a microphone array. The angle of the centerline of a camera field-of-view (FoV) and the angle of the camera FoV is determined. Knowing the angle from the centerline of the microphone array of the particular sound and then the angle of the centerline of the camera FoV and angles of the camera FoV allows a determination if the sound is inside the FoV of the camera. If so, the microphones are unmuted. If not, the microphones are muted. As the camera zooms or pans, the changes in camera FoV and centerline angle are computed and used with the sound angle, so that the muting and unmuting occurs automatically as the camera zoom and pan angle change.

CAMERA-VIEW ACOUSTIC FENCE
20230053202 · 2023-02-16 ·

Determining the angle of sound relative to the centerline of a microphone array. The angle of the centerline of a camera field-of-view (FoV) and the angle of the camera FoV is determined. Knowing the angle from the centerline of the microphone array of the particular sound and then the angle of the centerline of the camera FoV and angles of the camera FoV allows a determination if the sound is inside the FoV of the camera. If so, the microphones are unmuted. If not, the microphones are muted. As the camera zooms or pans, the changes in camera FoV and centerline angle are computed and used with the sound angle, so that the muting and unmuting occurs automatically as the camera zoom and pan angle change.

VIDEO COMMUNICATIONS APPARATUS AND METHOD
20230048798 · 2023-02-16 ·

Provided are apparatuses and associated methods for video communications and related features. In one embodiment, a big-screen video communications apparatus is provided that includes a projector and speaker for projecting received images and sounds and includes a camera and microphone for capturing images and sounds for transmission.

VIDEO COMMUNICATIONS APPARATUS AND METHOD
20230048798 · 2023-02-16 ·

Provided are apparatuses and associated methods for video communications and related features. In one embodiment, a big-screen video communications apparatus is provided that includes a projector and speaker for projecting received images and sounds and includes a camera and microphone for capturing images and sounds for transmission.

Creating a Printed Publication, an E-Book, and an Audio Book from a Single File
20230049537 · 2023-02-16 ·

As an example, a server may receive, from a computing device, a submission created by an author. The submission includes book data associated with a book and author data associated with the author. The author data includes incarceration data indicating whether the author was incarcerated. The server may determine, based on the author data and the book data, that the submission is publishable. The server may create, based on the book data, a printable book, an e-book, and an audio book and make one or more of the printable book, the e-book, and the audio book available for acquisition.

Creating a Printed Publication, an E-Book, and an Audio Book from a Single File
20230049537 · 2023-02-16 ·

As an example, a server may receive, from a computing device, a submission created by an author. The submission includes book data associated with a book and author data associated with the author. The author data includes incarceration data indicating whether the author was incarcerated. The server may determine, based on the author data and the book data, that the submission is publishable. The server may create, based on the book data, a printable book, an e-book, and an audio book and make one or more of the printable book, the e-book, and the audio book available for acquisition.

In-vehicle speech processing apparatus

An in-vehicle apparatus is connectable to a device that includes a voice assistant function. The in-vehicle apparatus includes: a voice detector that performs voice recognition of an audio signal input from a microphone and that controls functions of the in-vehicle apparatus based on a result of the voice recognition; and an interface that communicates with the device. When being informed of a detection of a predetermined word in the audio signal as the result of the voice recognition of the audio signal performed by the voice detector, the interface sends to the device, not via the voice detector, the audio signal input from the microphone. The predetermined word is for activating the voice assistant function of the device.

In-vehicle speech processing apparatus

An in-vehicle apparatus is connectable to a device that includes a voice assistant function. The in-vehicle apparatus includes: a voice detector that performs voice recognition of an audio signal input from a microphone and that controls functions of the in-vehicle apparatus based on a result of the voice recognition; and an interface that communicates with the device. When being informed of a detection of a predetermined word in the audio signal as the result of the voice recognition of the audio signal performed by the voice detector, the interface sends to the device, not via the voice detector, the audio signal input from the microphone. The predetermined word is for activating the voice assistant function of the device.

Pre-processing for automatic speech recognition

A method is provided that includes obtaining two or more microphone audio signals; analysing the two or more microphone audio signals for a defined noise type; and processing the two or more microphone audio signals based on the analysis to generate at least one audio signal suitable for automatic speech recognition. A corresponding apparatus is also provided.

ARTIFICIAL INTELLIGENCE-BASED AUDIO PROCESSING METHOD, APPARATUS, ELECTRONIC DEVICE, COMPUTER-READABLE STORAGE MEDIUM, AND COMPUTER PROGRAM PRODUCT
20230041256 · 2023-02-09 ·

An artificial intelligence-based audio processing method includes: obtaining an audio clip of an audio scene, the audio clip including noise; performing audio scene classification processing based on the audio clip to obtain an audio scene type corresponding to the noise in the audio clip; and determining a target audio processing mode corresponding to the audio scene type, and applying the target audio processing mode to the audio clip of the audio scene according to a degree of interference caused by the noise in the audio clip.