IP Library Granted Patent US 11,122,219
Granted Patent B2
US 11,122,219 · App. 16/542,383 · Granted Sep 14, 2021

Mid-span device and system

Inventors: John Zhang (San Jose, CA); Naveed Alam (Cupertino, CA); Yashket Gupta (San Jose, CA); Ram Natarajan (Cupertino, CA)
Assignee: ALTIA SYSTEMS INC.
H04N5/265G06F3/165H04N5/2628
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,122,219
App. No.
16/542,383
Granted
Sep 14, 2021
Kind
B2
Abstract

A video processing device which can be implemented as a peripheral device is disclosed. The video processing device is configured to receive one or more videos and audio through a plurality interfaces. Further, the device receives one or more user input corresponding to one or more customization requests of the received video and the audio. The video processing device processes the video and the audio based on the received user input and outputs the processed video and audio via an output interface.

Claims (26)

1. A peripheral video processing device, comprising:

a plurality of first interfaces to communicably connect the peripheral video device to a camera as a plug in, said first interfaces configured to receive one or more videos from the camera;

a second interface configured to receive audio from the camera;

a unit configured to split videos into multiple streams;

a pass-through processing unit configured to send the original stream as an output without any latency;

a module to receive one or more user inputs corresponding to one or more customization requests to customize the one or more videos and audio received;

one or more processing modules to make one or more customizations to one or more split videos and audio based on the one or more customization requests, wherein the one or more processing modules are configured to determine at least one of an active speaker in one or more received videos and a number of faces in the one or more received videos; and

an output interface to output customized video and audio.

2. The video processing device of claim 1 , wherein the one or more customizations comprises zooming in on the one or more videos.

3. The video processing device of claim 1 , wherein the one or more customizations comprises creating a composite video of the received one or more videos.

4. The video processing device of claim 1 , wherein the one or more videos are asynchronously received.

5. The video processing device of claim 1 , further comprising a plurality of buffers to store each of the received one or more videos.

6. The video processing device of claim 1 , wherein the one or more customizations comprises converting the received one or more videos into a H.264 video format.

7. The video processing device of claim 1 , wherein the one or more customizations comprises outputting a passthrough of the audio received.

8. The video processing device of claim 1 , wherein the output video corresponds to the video including the determined active speaker.

9. The video processing device of claim 1 , wherein the output interface outputs a plurality of videos simultaneously.

10. The video processing device of claim 1 , configured to recognize faces in the one or more received videos.

11. The video processing device of claim 10 , configured to recognize speech corresponding to the determined active speaker and generate one or more transcripts of the recognized speech.

12. The video processing device of claim 11 , wherein the generated one or more transcripts are associated with an ID of the active speaker.

13. A video processing method comprising:

receiving one or more videos through a plurality of first interfaces at a peripheral device plugged in to a camera;

receiving, at the peripheral device, an audio via a second interface;

processing, by the peripheral device, the one or more videos and the audio based on the received one or more user inputs comprising splitting the videos into multiple streams and sending, by a pass-through unit the original video stream as an output without any latency;

receiving, at the peripheral device, one or more user inputs corresponding to one or more customization requests corresponding to the one or more videos received and the audio received;

customizing, by the peripheral device, the one or more videos and the audio based on the received one or more user inputs comprising determining at least one of an active speaker in one or more received videos and a number of faces in the one or more received videos; and

outputting, by the peripheral device, the customized video and audio via an output interface.

Assignments (3)
MERGER Recorded Mar 30, 2026
From: GN AUDIO A/S
To: GN HEARING A/S
Reel/Frame 075299/0225 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 12, 2022
From: ALTIA SYSTEMS, INC.
To: GN AUDIO A/S
Reel/Frame 058995/0604 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2020
From: ZHANG, JOHN; ALAM, NAVEED; GUPTA, YASHKET; NATARAJAN, RAM
To: ALTIA SYSTEMS INC,
Reel/Frame 052531/0482 →