IP Library Granted Patent US 12,568,258
Granted Patent B2
US 12,568,258 · App. 18/503,858 · Granted Mar 3, 2026

Seamlessly inserting a supplemental content item into a content item

Inventors: Tao Chen (Palo Alto, CA); Reda Harb (Tampa, FL)
Assignee: Adeia Guides Inc.
H04N21/23424H04N21/23418H04N21/2393H04N21/262H04N21/44004
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,568,258
App. No.
18/503,858
Granted
Mar 3, 2026
Kind
B2
Abstract

The present disclosure relates to methods and systems, implemented by a device such as a client device, for seamlessly inserting a supplemental content item into a content item to negate the user's need for rewinding to a point prior to the interruption of the content item by the supplemental content item. The client device accesses the supplemental content insertion logic to identify a default supplemental content insertion point between two consecutive segments of the content item. The client device analyzes the two consecutive segments of the content item to identify a natural supplemental content insertion point within one of the two consecutive segments. The client device then decodes a first set of frames of the content item up to the natural supplemental content insertion point, a second set of frames of the supplemental content item and a third set of frames of the content item from the natural supplemental content insertion point. The client device places these three sets of frames in a buffer and plays the frames from the buffer.

Claims (73)

1 . A method comprising:

sending a request for a content item;

receiving at least a manifest for the content item;

accessing supplemental content insertion logic to identify a default supplemental content insertion point, based on the supplemental content insertion logic, between a first segment of the content item and a second segment of the content item, wherein the first segment and second segment are two consecutive segments;

analyzing the first segment of the content item and the second segment of the content item to identify a natural supplemental content insertion point within the first segment or the second segment;

in response to the identifying the natural supplemental content insertion point:

overriding insertion of a supplemental content item at the default supplemental content insertion point;

decoding a first set of frames of the content item up to the natural supplemental content insertion point and placing the first set of frames into a buffer;

decoding a second set of frames of the supplemental content item and placing the second set of frames into the buffer;

decoding a third set of frames of the content item, from the natural supplemental content insertion point, and placing the third set of frames into the buffer; and

playing frames from the buffer.

2 . The method of claim 1 , further comprising:

receiving a manifest for the supplemental content item;

receiving the first segment of the content item and the second segment of the content item using addresses provided by the manifest for the content item; and

receiving segments of the supplemental content item using addresses provided by the manifest for the supplemental content item.

3 . The method of claim 1 , wherein the identifying the natural supplemental content insertion point within the first segment or the second segment comprises:

identifying a portion of the first segment or the second segment, that does not comprise closed captions.

4 . The method of claim 1 , wherein the identifying the natural supplemental content insertion point within the first segment or the second segment comprises:

identifying a portion of the first segment or the second segment, that is associated with audio data of the content item that do not comprise speech.

5 . The method of claim 1 , wherein the supplemental content insertion logic is configured to identify the default supplemental content insertion point by setting a value corresponding to a number of segments of a sequence of segments of the content item intended to be played before starting playing segments of the supplemental content item, wherein the first segment is a last segment of the sequence of segments of the content item and the value corresponds to a place of the first segment in the sequence of segments of the content item.

6 . The method of claim 1 , wherein the identifying the natural supplemental content insertion point comprises:

identifying a plurality of natural supplemental content insertion points within any one of the first segment and the second segment; and

selecting a closest natural supplemental content insertion point from the plurality of natural supplemental content insertion points.

7 . The method of claim 1 , wherein the overriding insertion of the supplemental content item at the default supplemental content insertion point occurs in response to having the default supplemental content insertion point placed in between a boundary portion of the first segment and a boundary portion of the second segment, both the boundary portion of the first segment and the boundary portion of the second segment comprising closed captions.

8 . The method of claim 1 , wherein the overriding insertion of the supplemental content item at the default supplemental content insertion point occurs in response to having the default supplemental content insertion point placed in between a boundary portion of the first segment and a boundary portion of the second segment, both the boundary portion of the first segment and the boundary portion of the second segment being associated with audio data of the content item that comprise speech.

9 . The method of claim 1 , wherein the playing frames from the buffer comprises sequentially playing, from the buffer, the first set of frames, the second set of frames and the third set of frames.

10 . The method of claim 9 , wherein the sequentially playing, from the buffer, the first set of frames, the second set of frames and the third set of frames comprises:

playing audio data associated with the first set of frames while playing the first set of frames;

playing audio data associated with the second set of frames while playing the second set of frames; and

playing audio data associated with the third set of frames while playing the third set of frames.

11 . The method of claim 1 , wherein:

the decoding the first set of frames comprises decoding, by a first decoder, the first set of frames;

the decoding the second set of frames comprises decoding, by a second decoder, the second set of frames; and

the decoding the third set of frames comprises decoding, by the first decoder, the third set of frames; and

wherein the first decoder and the second decoder operate simultaneously.

12 . A system comprising:

input/output circuitry configured to:

send a request for a content item; and

receive at least a manifest for the content item; and

control circuitry configured to:

access supplemental content insertion logic to identify a default supplemental content insertion point, based on the supplemental content insertion logic, between a first segment of the content item and a second segment of the content item, wherein the first segment and second segment are two consecutive segments;

analyze the first segment of the content item and the second segment of the content item to identify a natural supplemental content insertion point within the first segment or the second segment; and

in response to the identifying the natural supplemental content insertion point:

override insertion of a supplemental content item at the default supplemental content insertion point;

decode a first set of frames of the content item up to the natural supplemental content insertion point and placing the first set of frames into a buffer;

decode a second set of frames of the supplemental content item and placing the second set of frames into the buffer;

decode a third set of frames of the content item, from the natural supplemental content insertion point, and placing the third set of frames into the buffer; and

play frames from the buffer.

13 . The system of claim 12 , wherein the input/output circuitry is further configured to:

receive a manifest for the supplemental content item;

receive the first segment of the content item and the second segment of the content item using addresses provided by the manifest for the content item; and

receive segments of the supplemental content item using addresses provided by the manifest for the supplemental content item.

14 . The system of claim 12 , wherein the control circuitry is configured to identify the natural supplemental content insertion point within the first segment or the second segment by:

identifying a portion of the first segment or the second segment, that does not comprise closed captions.

15 . The system of claim 12 , wherein the control circuitry is configured to identify the natural supplemental content insertion point within the first segment or the second segment by:

identifying a portion of the first segment or the second segment, that is associated with audio data of the content item that do not comprise speech.

16 . The system of claim 12 , wherein the supplemental content insertion logic is configured to identify the default supplemental content insertion point by setting a value corresponding to a number of segments of a sequence of segments of the content item intended to be played before starting playing segments of the supplemental content item, wherein the first segment is a last segment of the sequence of segments of the content item and the value corresponds to a place of the first segment in the sequence of segments of the content item.

17 . The system of claim 12 , wherein the control circuitry is configured to identify the natural supplemental content insertion point by:

identifying a plurality of natural supplemental content insertion points within any one of the first segment and the second segment; and

selecting a closest natural supplemental content insertion point from the plurality of natural supplemental content insertion points.

18 . The system of claim 12 , wherein the control circuitry is configured to override the insertion of the supplemental content item at the default supplemental content insertion point in response to having the default supplemental content insertion point placed in between a boundary portion of the first segment and a boundary portion of the second segment, both the boundary portion of the first segment and the boundary portion of the second segment being associated with audio data of the content item that comprise speech.

19 . The system of claim 12 , wherein the control circuitry is configured to play frames from the buffer, by sequentially playing, from the buffer, the first set of frames, the second set of frames and the third set of frames.

20 . A system comprising:

means for sending a request for a content item;

means for receiving at least a manifest for the content item;

means for accessing supplemental content insertion logic to identify a default supplemental content insertion point, based on the supplemental content insertion logic, between a first segment of the content item and a second segment of the content item, wherein the first segment and second segment are two consecutive segments;

means for analyzing the first segment of the content item and the second segment of the content item to identify a natural supplemental content insertion point within the first segment or the second segment; and

means for, in response to the identifying the natural supplemental content insertion point:

overriding insertion of a supplemental content item at the default supplemental content insertion point;

decoding a first set of frames of the content item up to the natural supplemental content insertion point and placing the first set of frames into a buffer;

decoding a second set of frames of the supplemental content item and placing the second set of frames into the buffer;

decoding a third set of frames of the content item, from the natural supplemental content insertion point, and placing the third set of frames into the buffer; and

playing frames from the buffer.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 16, 2024
From: CHEN, TAO; HARB, REDA
To: ADEIA GUIDES INC.
Reel/Frame 066135/0745 →
Continuity (1)
Related Publication 20250150651A1 · May 8, 2025
References Cited (15)
US 8347344B2 · Makhija · 2013 [cited by examiner]
US 10880585B1 · Waggoner · 2020 [cited by examiner]
US 11245935B1 · Shams et al. · 2022 [cited by applicant]
US 20140365675A1 · Bhardwaj et al. · 2014 [cited by applicant]
US 20200413139A1 · Ickin · 2020 [cited by applicant]
US 20210014542A1 · Jimenez et al. · 2021 [cited by applicant]
US 20220101013A1 · Chatoo · 2022 [cited by examiner]
YouTube's lowered mid-roll ad requirement may lead to shorter videos from publishers—Digiday by Tim Peterson Jul. 24, 2020. [cited by applicant]
YouTube Threatens to Cut off Ad Blocker Users After Just Three Ad-less Vids (gizmodo.com) by Kyle Barr, published Jun. 30, 2023. [cited by applicant]
Manage mid-roll ad breaks in long videos—YouTube Help (google.com) Printed on Dec. 5, 2023. [cited by applicant]
Z, Liu, “A deep neural framework to detect individual advertisement (ad) from videos,” in Proceedings of IEEE/CVF WACV, 2023. [cited by applicant]
R. Tapu, et al., “Deep-Ad: A Multimodal Temporal Video Segmentation Framework for Online Video Advertising,” IEEE Access, May 2020. [cited by applicant]
MPEG-4, Advanced Video Coding (Part 10) (H.264), H.264: Advanced video coding for generic audiovisual services (itu.int) Approved in Aug. 22, 2021. [cited by applicant]
MPEG-H Part 2, High Efficiency Video Coding H.265 : High efficiency video coding (itu.int) Approved in Aug. 22, 2021. [cited by applicant]
Thomas Stockhammer (Qualcomm): “[Dash] Content Replacement and Ad Insertion Event”, 126. MPEG Meeting; Mar. 25, 2019-Mar. 29, 2019; Geneva; (Motion Picture Expert Group or ISO/IEC JTC1/SC29/WG11), No. m47557, Mar. 20, 2… [cited by applicant]