Zero-UI here means a product that acts on its own and has little screen to show, and interface-free narrative is a coded class, not a judgement. Across 47 autonomous and agentic AI product videos from 37 companies, uploaded 12 January 2026 to 3 August 2026, 42 of 47 still show a product interface and 7 go interface-free: Atlassian, Celigo, Genpact, GitLab, etc. Dashboards appear in 33 videos and chat panels in 24, kinetic type stands in for the screen in 27, and 8 of the 16 screen-led videos drop the voiceover, 5 of them from Anthropic. The set is thin: 47 videos, under this series' floor of 80.
ADVIDS Research · 47 autonomous AI videos coded · 37 companies · 423 sampled frames · 496 timed segments · published 14 September 2026 · 18 min read
Top takeaways
Autonomous AI videos keep the product screen; seven go interface-free
42 of 47 autonomous AI product videos uploaded January to August 2026 show a product interface in at least one segment; 7 are interface-free by this report's definition: Atlassian, Celigo, Genpact, GitLab, etc.
Dashboards, not chat panels, carry the screen-led AI agent video
Dashboards appear in 33 of 47 autonomous AI videos from January to August 2026 and chat panels in 24; of the 16 videos built at least half on screens, 8 are dashboard-led and 4 chat-led (Anthropic, Perplexity, Zoom and OpenAI).
Half the screen-led videos drop the voiceover, Anthropic's five included
8 of the 16 screen-led autonomous AI videos from January to August 2026 run music only, with no voiceover, and 5 of them are Anthropic's; outside Anthropic and OpenAI, 1 of 9 do.
Kinetic type is the stand-in autonomous AI videos reach for first
Kinetic type appears in 27 of 47 autonomous AI videos from January to August 2026 and is the first interface-free motif in 23 of the 38 that carry one; Dext, IPRally and Infobip build around it, and robot hardware appears in none.
The interface-free layer opens; the product screen arrives second
In 24 of the 34 autonomous AI videos from January to August 2026 that carry both, an interface-free motif comes before the first product screen; SAP, Anthropic, Asana, Campfire, etc. show the screen first.
Seven interface-free videos run on drawn 2D, type and voiceover
The 7 interface-free autonomous AI videos from January to August 2026 are built with 2D animation in 6 cases and narrated in 5: Celigo, GitLab, IBM, Salesforce, etc. carry a voiceover, while Genpact and Atlassian run on music.
Questions this report answers
How do AI agent videos show a product with no interface?
7 of 47 autonomous AI product videos go interface-free, with product screens under a tenth of runtime and no interface in the coded surface. They lean on kinetic type, flowcharts, network diagrams and b-roll; companies include Atlassian, Celigo, Genpact, GitLab, etc.
Do agentic AI videos still show the product UI?
Yes: 42 of 47 show a dashboard, chat panel, terminal, device screen or screen capture in at least one segment. The median video keeps product screens on for 43 per cent of its runtime, and 16 build at least half of it on screens.
Should an AI agent video use a chat interface or a dashboard?
The set leans to dashboards: 33 of 47 videos show one against 24 with a chat panel, and where they appear dashboards cover a median 25 per cent of runtime against 21 for chat. Among the 16 screen-led videos, 8 are dashboard-led and 4 chat-led.
What visuals replace the interface in autonomous AI videos?
Kinetic type first: 27 of 47 videos carry it, at a median 0:04 into the video. Flowcharts or network diagrams appear in 20, notifications in 6, abstract AI orbs in 4, live-action b-roll in 4 and robot hardware in none.
Do AI agent launch videos use voiceover?
28 of 47 videos carry a voiceover and 19 run music only. 5 of the 7 interface-free videos are narrated, while 8 of the 16 screen-led videos run without one, every screen-led video from Anthropic and OpenAI among them.
How was this analysis done?
ADVIDS Research screened 188 visually coded product videos against a rule written before intake, and 47 from 37 companies qualified. Each was coded from nine sampled frames and timed segments, and the coded dataset ships with the report.
How to read this report
- The set is 47 autonomous and agentic AI product videos from 37 companies, uploaded to YouTube 12 January 2026 to 3 August 2026. It is one catalogue with no periods, so nothing here describes change over time.
- “N of M” counts videos: N carry the coded value and M is the base named beside it. A base under 20 is written as a count, and 47 is under this series' floor of 80, so every share describes these videos only.
- A video is interface-free when product screens cover under a tenth of runtime and the coded surface lists no interface, screen-led when screens cover half or more, and blended otherwise.
Finding 01 · Screen · Screens on · 47 autonomous AI videos, January to August 2026
1. Autonomous AI videos keep the product screen; seven go interface-free
42 of 47 autonomous and agentic AI product videos uploaded between 12 January 2026 and 3 August 2026 show a product screen in at least one coded segment, and 35 list an interface in the coded surface. Product screens cover a median 43 per cent of runtime across the set and a quarter or more in 34 videos. Only 7 meet this report's interface-free definition, and 5 carry no screen motif at all. The agent is sold as acting on its own, and the video still shows where it acts.
Companies in this group include Anthropic, SAP, OpenAI, Workday, etc. Without Anthropic, the largest company in the pattern, the share moves 1.3 points.
Swimlane goes the other way. Its February 2026 Symphony of AI Agents spot runs 0:48 with no product screen: kinetic-type lines cover 73 per cent of runtime around a drum kit and a silhouetted analyst, under a 110-word voiceover. Where the agents work inside a security team's existing tools, the asset shows the analyst's role rather than a screen.
Source: ADVIDS Research coded set, 47 videos from 37 companies, uploaded 12 January 2026 to 3 August 2026

The interface-free video is the exception this catalogue has to explain, not the default of the category.
Finding 02 · Screen · Dashboard or chat · 47 autonomous AI videos, January to August 2026
2. Dashboards, not chat panels, carry the screen-led AI agent video
Dashboards appear in 33 of 47 autonomous AI videos from January to August 2026 and chat panels in 24, a gap of 9 videos. Where they appear, dashboards cover a median 25 per cent of runtime against 21 per cent for chat, and a dashboard covers a quarter or more of the video in 18 against 8 for chat. Of the 16 screen-led videos, 8 are dashboard-led, 4 chat-led and 4 led by a terminal or screen capture. The chat panel arrives earlier, at a median 0:16 against 0:25. Chat arrives first; the dashboard stays on screen longer.
Companies in this group include Alloy, Amplience, Asana, Savant Labs, etc. Without SAP, the dashboard share moves 2.0 points.
Zoom goes the other way. Its March 2026 Zoom Virtual Agent tour runs 2:11, and chat panels cover 34 per cent of runtime against 25 per cent for the agent-management dashboard. Where the agent is itself the support conversation, the chat panel is the product rather than a stand-in for it.
Source: ADVIDS Research coded set, 47 videos from 37 companies, uploaded 12 January 2026 to 3 August 2026

The chat panel is not the default picture of an agent; the view of what the agent did is.
Finding 03 · Sound and type · Voiceover · 47 autonomous AI videos, January to August 2026
3. Half the screen-led videos drop the voiceover, Anthropic's five included
8 of the 16 autonomous AI videos built at least half on product screens run music only, with no voiceover, and the median screen-led video carries 8 spoken words against 148 for the interface-free class and 140 for blended videos. The split belongs to two companies: all 7 screen-led videos from Anthropic and OpenAI run on music, 5 of them from Anthropic, against 1 of the other 9. Without Anthropic the music-only share falls 22.7 points, to 3 of 11. At the other end, 5 of the 7 interface-free videos and 15 of the 24 blended videos are narrated. With no screen, a voice carries the product; with a screen, two companies let the interface run silent.
Companies in this group include Anthropic, OpenAI and Make.
Blink goes the other way. Its January 2026 Agent Builder video runs 0:55 on screen capture and title cards, with product screens covering 65 per cent of runtime, and carries a 67-word voiceover over the build. Where the viewer has to follow a prompt becoming an app, the voice names each step the screen shows.
Source: ADVIDS Research coded set, 47 videos from 37 companies, uploaded 12 January 2026 to 3 August 2026

The silence belongs to two companies: Anthropic and OpenAI cut every screen-led video to music, and the other vendors narrate.
Finding 04 · Sound and type · Kinetic type · 47 autonomous AI videos, January to August 2026
4. Kinetic type is the stand-in autonomous AI videos reach for first
Kinetic type appears in 27 of 47 autonomous AI videos from January to August 2026, arrives at a median 0:04, and covers a median 30 per cent of runtime where it appears. It is the first interface-free motif in 23 of the 38 videos that carry one, and covers a quarter or more of runtime in 16. Flowcharts and network diagrams appear in 12 and 12 videos, 20 with either. By class, kinetic typography runs in 5 of the 7 interface-free videos and 7 of the 16 screen-led ones. Before a diagram, an orb or a robot, the set writes the agent's work out as words on screen.
Companies in this group include Atlassian, Board, Celigo, Contentstack, etc. Without Anthropic, the kinetic-type share moves 0.3 points.
Genpact goes the other way. Its April 2026 accounts-payable spot runs 0:26 with no kinetic type and no product screen: live-action b-roll of an orchestra whose musicians have invoice-paper heads covers 84 per cent of runtime, on music alone. Where the claim is a metaphor for autonomous processing, the picture carries it without a written line.
Source: ADVIDS Research coded set, 47 videos from 37 companies, uploaded 12 January 2026 to 3 August 2026

Type leads the next interface-free motifs, flowcharts and network diagrams in 12 videos each, by 15 videos.
Finding 05 · Order · Order · 47 autonomous AI videos, January to August 2026
5. The interface-free layer opens; the product screen arrives second
In 24 of the 34 autonomous AI videos from January to August 2026 that carry both a product screen and an interface-free motif, the interface-free motif comes first; in 10 the screen does. Across the set the first interface-free motif lands at a median 0:05 and the first product screen at 0:12. 22 of 47 videos open on a title or logo card and 5 on a product screen, and 40 of 47 close on a card. In the screen-led class the order reverses: the first screen lands at a median 0:04 and the first interface-free motif at 0:08, among the 9 that carry one. The words and diagrams come first; the screen where the agent works comes second.
Companies in this group include Dext, ElevenLabs, Genpact, GitLab, etc. Without Anthropic, the share moves 0.4 points.
Campfire goes the other way. Its March 2026 Ember Agents video runs 0:50: after a title card, the accounting dashboard arrives at 0:04 and covers 46 per cent of runtime, and the only interface-free motif, a notification, waits until 0:39. Where the agents' output is a queue of items to approve, the queue opens the video.
Source: ADVIDS Research coded set, 47 videos from 37 companies, uploaded 12 January 2026 to 3 August 2026

The first product screen lands at a median 0:12; the opening seconds go to type, diagrams and cards.
Finding 06 · Build · Interface-free · 47 autonomous AI videos, January to August 2026
6. Seven interface-free videos run on drawn 2D, type and voiceover
7 of 47 autonomous AI videos from January to August 2026 meet the interface-free definition. 6 of the 7 are built with 2D animation, 5 read as motion graphics, 5 carry a voiceover and 5 use kinetic type; flowcharts or network diagrams appear in 4. Interface-free motifs cover a median 71 per cent of their runtime, and title or logo cards 29 per cent against 14 for screen-led videos; 6 of the 7 open on a card. Without a screen, the product becomes a drawn system: nodes, lines of type and a narrator.
Companies in this group include IBM, Salesforce, Swimlane, etc.
Salesforce goes the other way inside the class. Its April 2026 Agentic Orchestration explainer runs 1:31 with no product screen and none of the seven interface-free motifs: a photo collage of a jazz band covers 65 per cent of runtime between title cards, under a 220-word voiceover that casts the orchestrator as the bandleader. Where the product is coordination between agents, a metaphor stands in for both the screen and the diagram.
Source: ADVIDS Research coded set, 47 videos from 37 companies, uploaded 12 January 2026 to 3 August 2026

None of the seven runs on light frames: median brightness is 79 against 190 for screen-led videos, on a 0 to 255 scale.
Benchmarks by strategy class
Descriptive figures for placing your own autonomous AI video against each class.
Screens, stand-ins, sound and length
| Measure | Interface-free (7) | Blended (24) | Screen-led (16) | All (47) | What it means |
|---|---|---|---|---|---|
| Runtime, median | 1:14 | 1:17 | 1:03 | 1:14 | Half the videos in the class run shorter than this |
| Product-screen coverage, median per cent | 0 | 37 | 68 | 43 | Per cent of runtime with a screen motif on screen |
| Interface-free motif coverage, median per cent | 71 | 35 | 8 | 25 | Per cent of runtime with kinetic type, a diagram, an orb, b-roll, robot hardware or a notification |
| Title and logo card coverage, median per cent | 29 | 12 | 14 | 15 | Per cent of runtime on a title or logo card |
| First product screen, median | 0:44 | 0:14 | 0:04 | 0:12 | Among videos that show one: interface-free 3, blended 23, screen-led 16 |
| First interface-free motif, median | 0:04 | 0:06 | 0:08 | 0:05 | Among videos that carry one: interface-free 6, blended 23, screen-led 9 |
| Words spoken, median | 148 | 140 | 8 | 115 | Words in the captions; music-only videos with no captions count as zero |
| Frame brightness, median | 79 | 155 | 190 | 155 | Mean luminance across the nine sampled frames, 0 to 255 |
What most autonomous AI videos skip
| Coded value | Videos | Companies |
|---|---|---|
| Robot hardware on screen | 0 of 47 | none |
| Handheld device screen | 3 of 47 | Contentstack, Ofelia and Workday |
| AI orb or abstract AI shape | 4 of 47 | ElevenLabs, IBM, Nexthink and Nooks |
| Live-action b-roll | 4 of 47 | Genpact, SAP, StackAI and Swimlane |
| Presenter on camera | 4 of 47 | Nooks, Ofelia, SAP and StackAI |
What to do with this
- Finding 01 If 42 of 47 autonomous AI videos show a product screen, budget for a presentable demo environment even when the product acts in the background.
- Finding 02 If dashboards appear in 33 of 47 and chat panels in 24, design the screen that shows what the agent did before the chat that shows what it was asked.
- Finding 03 If 1 of the 9 screen-led videos outside Anthropic and OpenAI dropped the voiceover, keep the narration line in a screen-led brief unless you are cutting to that pair's pattern on purpose.
- Finding 04 If kinetic type appears in 27 of 47, write the on-screen lines with the script, not after the edit.
- Finding 05 If the first product screen lands at a median 0:12, storyboard the opening seconds as type or a diagram and schedule the demo environment for the middle of the video.
- Finding 06 If 6 of the 7 interface-free videos are 2D animation, brief a 2D studio and a narrator when there is no screen to show.
Methodology
| Vertical | Videos |
|---|---|
| devtools | 9 |
| fintech | 7 |
| cx | 4 |
| productivity | 3 |
| ai | 3 |
| enterprise | 2 |
| other verticals | 19 |
| Coding scheme | Values | Definition |
|---|---|---|
| Strategy class | interface-free · blended · interface fallback (screen-led) | Computed from coded segments: interface-free when product-screen motifs cover under 0.10 of runtime and the coded surface lists no dashboard, chat or terminal; interface fallback when they cover 0.50 or more; blended otherwise. Fallback videos are sub-typed chat-led, dashboard-led or other screen-led by the screen motif with the most runtime. |
| Screen motif | dashboard_ui · chat_ui · terminal_ui · handheld_device_ui · screen_recording_generic | Coded per timed segment from frames, with the first timecode and the share of runtime each motif covers. |
| Interface-free motif | kinetic_type · diagram_flowchart · diagram_network · ai_orb_abstract · cinematic_broll · robot_hardware · notification_alert | Coded per timed segment from frames. cinematic_broll is live-action b-roll; notification_alert is a toast or alert card. |
| Surface | dashboard · chat · terminal · diagram · none | Coded from frames: the interface a video shows. A video can list several. |
| Cards | title_card · brand_logo_card | Coded per timed segment. The first and last segments give the opening and closing frame. |
| Build and format | 2d_animation · ui_mockup · screen_recording · kinetic_type · live_action · 3d_cgi · screencast · motion_graphics | Coded from frames with a confidence of high or medium. A video can carry several. |
| Sound and frame | voiceover · music_only · word count · brightness 0 to 255 | Voiceover or music only from audio analysis; words from captions; brightness computed over the nine sampled frames, not the interface alone. |
| Not coded | cast · altitude | Who appears and what the video argues. The library records neither, and this report does not code them from transcripts. |
Cite this report
Jai Ghosh, Advids. Decoding the Zero-UI Challenge: 47 Autonomous AI Videos Audited for Interface-Free Narrative. 14 September 2026. https://advids.co/blog/zero-ui-autonomous-ai-video-narrative
Author & editor bio
About Advids
Advids produces video for AI, SaaS and DeepTech companies. ADVIDS Research publishes coded studies of what those companies ship, with the dataset attached so every count can be checked.
Image credits and licensing
Video frames on this page are reproduced for analysis and commentary. Copyright in each frame remains with the company credited in its caption; to license a frame, contact that rights holder. Advids grants no rights over them.
Charts and images credited ADVIDS Research, and the coded dataset behind them, are © 2026 Advids and licensed CC BY 4.0: reuse them with attribution to Advids and a link to this page.