<?xml version="1.0" encoding="UTF-8" standalone="no"?><rss xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" version="2.0">
  <channel>
    <title>VP Land (FKA Coffee &amp; Celluloid)</title>
    <description>News, trends, and insights on the latest creative technology</description>
    
    <link>https://newsletter.vp-land.com/</link>
    <atom:link href="https://rss.beehiiv.com/feeds/pQ7FTxHZ8E.xml" rel="self"/>
    
    <lastBuildDate>Thu, 20 Aug 2026 03:22:18 +0000</lastBuildDate>
    <pubDate>Tue, 21 Jul 2026 18:51:09 +0000</pubDate>
    <atom:published>2026-07-21T18:51:09Z</atom:published>
    <atom:updated>2026-08-20T03:22:18Z</atom:updated>
    
      <category>Film</category>
      <category>Media</category>
      <category>Technology</category>
    <copyright>Creative Commons Attribution-Share Alike</copyright>
    
    <image>
      <url>https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/publication/logo/cdda2bd5-66ce-4e18-9841-f1a2bb3c031e/VP_Land_Avatar_2025_-_10.jpg</url>
      <title>VP Land</title>
      <link>https://newsletter.vp-land.com/</link>
    </image>
    
    <docs>https://www.rssboard.org/rss-specification</docs>
    <generator>beehiiv</generator>
    <language>en-us</language>
    <webMaster>support@beehiiv.com (Beehiiv Support)</webMaster>

      <itunes:explicit>no</itunes:explicit><itunes:image href="https://newterritory.media/wp-content/uploads/2021/01/CC-Podcast-1.png"/><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords><itunes:summary>Coffee and Celluloid is a blog and podcast exploring film, filmmaking, and new media, looking at both the final product and the process of getting there. &#13;
&#13;
This podcast features interviews, Q&amp;As, and conversations with independent filmmakers.&#13;
&#13;
Hosted by Joey Daoud, Cherie Saulter, Carlos Rivera, and Andrew Hevia.</itunes:summary><itunes:subtitle>Freshly brewed podcast on film, photography, and all things visual</itunes:subtitle><itunes:category text="TV &amp; Film"/><itunes:category text="Arts"><itunes:category text="Visual Arts"/></itunes:category><itunes:category text="Society &amp; Culture"><itunes:category text="Places &amp; Travel"/></itunes:category><itunes:category text="Society &amp; Culture"><itunes:category text="Personal Journals"/></itunes:category><itunes:category text="Arts"><itunes:category text="Design"/></itunes:category><itunes:author>Coffee and Celluloid</itunes:author><itunes:owner><itunes:email>podcast@newterritory.media</itunes:email><itunes:name>Coffee and Celluloid</itunes:name></itunes:owner><item>
  <title>NVIDIA puts AI agents inside your creative tools</title>
  <description>PLUS: Alibaba's 2.4T model, Google's free head rig, a Resolve agent</description>
      <enclosure length="985504" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b6fdd8b4-e09b-45e2-9d08-311ec37311c3/00-hero-banner__1_.jpg"/>
  <link>https://newsletter.vp-land.com/p/nvidia-puts-ai-agents-inside-your-creative-tools</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/nvidia-puts-ai-agents-inside-your-creative-tools</guid>
  <pubDate>Tue, 21 Jul 2026 18:51:09 +0000</pubDate>
  <atom:published>2026-07-21T18:51:09Z</atom:published>
    <category><![CDATA[Newsletter]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><hr class="content_break"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0851b96e-9574-41fd-a066-edbf06c8844c/VP_Land_Newsletter_Banner_-_18.png"/></div><hr class="content_break"><div class="section" style="background-color:#FFFFFF;border-color:#c75a5e;border-radius:50px;border-style:solid;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;"><b>Welcome to VP Land! </b>Open source keeps reaching deeper into production, from calibrating LED stages to modeling a digital human&#39;s face.</p><p class="paragraph" style="text-align:left;">Two for the calendar: PIXERA&#39;s <a class="link" href="https://www.vp-land.com/p/hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-se?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">HDR and Digital Color in Realtime Video Workflows</a> runs this Friday in Santa Monica, with a free YouTube livestream. And for SIGGRAPH, we have the <a class="link" href="https://www.eventbrite.com/e/ai-workflows-summit-tickets-1993363191967?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">AI Workflows Summit</a>, a free event this evening. I&#39;ll be moderating a panel on hybrid film production workflows with AI.</p><p class="paragraph" style="text-align:left;">In today&#39;s edition:</p><ul><li><p class="paragraph" style="text-align:left;">ASC open-sources an LED-volume test kit</p></li><li><p class="paragraph" style="text-align:left;">NVIDIA puts local AI agents across creative tools</p></li><li><p class="paragraph" style="text-align:left;">Alibaba previews a 2.4T-parameter model</p></li><li><p class="paragraph" style="text-align:left;">Google frees a scan-based head model</p></li></ul></div><hr class="content_break"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1a8ef6ca-1f50-47f7-bb27-bde9ab288e6c/Video_Tape_-_Box_Small.png?t=1712408310"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-width:0px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">The ASC Is Open-Sourcing a Virtual Production Test Kit for LED Volumes</h3></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e7891423-ac1c-419e-a76e-67ad599edc52/asc-led-test-kit.jpg?t=1784653570"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">The American Society of Cinematographers is open-sourcing <b>StEM3-VP</b>, a set of reference assets for evaluating and calibrating LED volumes, and the material is joining the Academy Software Foundation&#39;s <a class="link" href="https://dpel.aswf.io/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Digital Production Example Library</a>.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Built for the LED-to-camera seam.</b> The assets test the image path between the wall and the camera, where color, moiré, and brightness problems surface on a volume. A production can characterize and sign off a stage before principal photography, using open media instead of vendor-specific test patterns.</p></li><li><p class="paragraph" style="text-align:left;"><b>Openly licensed, no vendor lock.</b> StEM3-VP sits in DPEL alongside other production-grade test content, giving cinematographers, DITs, and stage engineers a fixed target that does not belong to a single wall or processor manufacturer.</p></li><li><p class="paragraph" style="text-align:left;"><b>A two-decade lineage.</b> It follows the ASC&#39;s original StEM from 2004 and StEM2 in 2022, extending that evaluation approach from digital projection and color pipelines to the LED stage.</p></li><li><p class="paragraph" style="text-align:left;"><b>Surfaced through ASWF&#39;s open program.</b> We covered the <a class="link" href="https://www.vp-land.com/p/dreamworks-moonray-keynote-headlines-aswf-open-source-days-in-la?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Open Source Days schedule</a> where StEM3-VP was slated to present, on a track spanning color-fidelity standards, OpenPBR, and studio-safe AI workflows.</p></li></ul></div><hr class="content_break"><div class="section" style="background-color:transparent;border-color:#3a8acc;border-radius:50px;border-style:solid;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:center;"><span style="color:#0194f6;"><b>SPONSOR MESSAGE</b></span></p><h3 class="heading" style="text-align:left;">The best prompt engineers aren&#39;t typing. They&#39;re talking.</h3><div class="image"><a class="image__link" href="https://ref.wisprflow.ai/beehiiv-ai/?utm_campaign={{publication_alphanumeric_id}}&utm_source=beehiiv&utm_term=ai_p4_q3&_bhiiv=opp_46683432-da1d-4766-8421-d34008df8a05_4de8c0ec&bhcl_id=4c6eb038-02fb-4ffb-b414-98418ac7ced9_{{subscriber_id}}_{{email_address_id}}" rel="noopener" target="_blank"><img class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9edd18e6-e2c3-47f7-9315-469682fd5892/flow-top-teams-move-faster.png?t=1776897861"/></a></div><p class="paragraph" style="text-align:left;">Power users figured this out early: speaking a prompt gives you 10x more context in half the time. You include the edge cases, the examples, the tone you want — because talking is fast enough that you don&#39;t skip them.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://ref.wisprflow.ai/beehiiv-ai/?utm_campaign={{publication_alphanumeric_id}}&utm_source=beehiiv&utm_term=ai_p4_q3&_bhiiv=opp_46683432-da1d-4766-8421-d34008df8a05_4de8c0ec&bhcl_id=4c6eb038-02fb-4ffb-b414-98418ac7ced9_{{subscriber_id}}_{{email_address_id}}" target="_blank" rel="noopener noreferrer nofollow">Wispr Flow</a> captures everything you say and turns it into clean, structured text for any AI tool. Speak messy. Get polished input. Paste into ChatGPT, Claude, Cursor, or wherever you work.</p><p class="paragraph" style="text-align:left;">89% of messages sent with zero edits. 4x faster than typing. Works system-wide on Mac, Windows, and iPhone.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://ref.wisprflow.ai/beehiiv-ai/?utm_campaign={{publication_alphanumeric_id}}&utm_source=beehiiv&utm_term=ai_p4_q3&_bhiiv=opp_46683432-da1d-4766-8421-d34008df8a05_4de8c0ec&bhcl_id=4c6eb038-02fb-4ffb-b414-98418ac7ced9_{{subscriber_id}}_{{email_address_id}}" target="_blank" rel="noopener noreferrer nofollow">Start flowing free</a></p></div><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><hr class="content_break"></div><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/845ebb56-17b5-4d6f-9526-5e36c53ed637/Light_-_Box_Small.png?t=1720210827"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-width:0px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">NVIDIA Launches Local AI Agents and More</h2></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/cfd6a121-fb54-4695-96fa-fc86a0981895/nvidia-local-agents.png?t=1784653593"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">NVIDIA used SIGGRAPH to move AI agents directly into the tools artists already use, with a cluster of releases aimed squarely at creative and production work.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Local agents in Blender.</b> An open <a class="link" href="https://blogs.nvidia.com/blog/siggraph-news-2026/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools#nemoclaw-dgx-station" target="_blank" rel="noopener noreferrer nofollow">Agent Toolkit</a> lets teams run AI agents inside Blender on their own hardware, keeping unreleased assets, scripts, and client material on-site instead of routing them through a third-party API.</p></li><li><p class="paragraph" style="text-align:left;"><b>Agents across more creative apps.</b> NVIDIA <a class="link" href="https://blogs.nvidia.com/blog/siggraph-news-2026/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools#mcp" target="_blank" rel="noopener noreferrer nofollow">lined up a wave of creative software behind MCP</a>, the open standard that lets an agent read and act on an application&#39;s live state, extending the same automation well beyond a single tool.</p></li><li><p class="paragraph" style="text-align:left;"><b>A synthetic-video detector for newsrooms.</b> A new <a class="link" href="https://blogs.nvidia.com/blog/siggraph-news-2026/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools#synthetic-video" target="_blank" rel="noopener noreferrer nofollow">Synthetic Video Detector microservice</a> flags whether a clip contains synthetic content, a sign that provenance is becoming part of the same creative stack.</p></li></ul></div><hr class="content_break"><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:0px;border-bottom-right-radius:0px;border-bottom-width:0px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:50px;border-top-right-radius:50px;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">Alibaba Previews Qwen3.8-Max, a 2.4-Trillion-Parameter Multimodal Model With Open Weights to Follow</h2></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/947e4c7a-d1cb-4276-bf90-5398edae410f/qwen-3-8-max.jpg?t=1784653605"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">Alibaba&#39;s Qwen team opened preview access to <b>Qwen3.8-Max</b>, a 2.4-trillion-parameter multimodal model it calls its most capable yet, and the latest sign that China&#39;s open-weight labs are shipping frontier-class systems right behind Kimi and Fable.</p><ul><li><p class="paragraph" style="text-align:left;"><b>In paid preview now.</b> It runs through Alibaba&#39;s Token Plan, Qoder, and QoderWork at 10 percent of standard pricing and takes images, video, and documents as input, with open weights promised at <a class="link" href="https://x.com/alibaba_qwen/status/2078759124914098291?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">a later date</a>.</p></li><li><p class="paragraph" style="text-align:left;"><b>The ranking is a vendor claim.</b> Qwen says it trails only Fable 5 among frontier systems, but it has published no benchmarks and has not disclosed the active-parameter count or its mixture-of-experts setup, so 2.4 trillion is total size, not per-query compute.</p></li></ul></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e619d014-9e24-4878-b569-04b3b2d62bb8/Toolbox_-_Box_Small.png?t=1738938789"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-width:0px;border-color:#3a8acc;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">Google Open-Sources GNM Head, a Parametric 3D Head Model With 250+ Identity Controls</h3></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a2a6e858-c7be-48fa-8e91-a432d17c621f/gnm-head.jpg?t=1784653625"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#3a8acc;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">Google open-sourced <b>GNM Head</b>, a <a class="link" href="https://www.vp-land.com/p/google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">scan-based parametric 3D head model</a>, under a commercial-friendly Apache 2.0 license, moving a studio-grade building block out of the research lab and into production pipelines. The model exposes 636 controls spanning identity and expression, down to fine detail in the eyes, teeth, and tongue, so you can dial in a specific face or drive a performance without hand-sculpting every blendshape. A free Blender 5.1+ importer ships alongside it, turning those controls into live sliders in the viewport with no plugin fees or research-only strings attached. For anyone rigging digital humans, populating background crowds, or building avatar systems, it is a rare genuinely production-ready open asset: commercially licensed, usable today in a tool most artists already run, and detailed enough to hold up in close-up work.</p></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7b3e63fb-94ff-4fda-9feb-6daa731cfb93/Television_-_Box_Small.png?t=1738625951"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#b4d8db;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">Neill Blomkamp’s NIGHTBORNE: a 13-minute Seedance 2.0 AI-film test</h3><p class="paragraph" style="text-align:left;">Neill Blomkamp’s 13-minute sci-fi horror film is the first release from Barley Studios, set in Peter Watts’ <i>Echopraxia</i> universe. Its production uses Seedance 2.0 to generate every frame, with Blomkamp directing shot by shot through prompts.</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/8Wbtt2JxP7g" width="100%"></iframe><p class="paragraph" style="text-align:left;">The “fully AI” label needs context: the project also uses real concept artists and licensed faces and voices from 32 human performers. That makes it a useful watch for the emerging hybrid model, where generative imagery carries the shots while artists and performers still supply core creative inputs.</p></div><hr class="content_break"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7d8b5c9f-853b-47b9-a096-82a3f91451a6/Clothespin_-_Box_Small.png?t=1738623074"/></div><div class="section" style="background-color:transparent;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#c40101;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;"><i>Stories, projects, and links that caught our attention from around the web:</i></p><p class="paragraph" style="text-align:left;">▶️ Alibaba&#39;s <a class="link" href="https://x.com/Alibaba_Wan/status/2075463287630876817?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Wan-Streamer</a> runs real-time, full-duplex video conversation at about 550ms and can embody any character.</p><p class="paragraph" style="text-align:left;">🎬 AI video startup Invideo is making a feature film and <a class="link" href="https://x.com/invideoOfficial/status/2075633658338423137?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">soliciting scripts</a> at <a class="link" href="mailto:made@invideo.io" target="_blank" rel="noopener noreferrer nofollow">made@invideo.io</a>.</p><p class="paragraph" style="text-align:left;">👩🏻‍💻 A free, open-source <a class="link" href="https://x.com/SamJWasserman/status/2076320682636439699?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">DaVinci Resolve MCP</a> hands an AI agent 37 live tools across color, sound, and timeline assembly.</p></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2bd899e4-ee40-4667-8548-4ba64b6fef8f/Denoised_-_Box_Small_-_01.png?t=1745256407"/></div></div><div class="section" style="background-color:transparent;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#7600c3;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">We Watched Codex Rebuild a Real Theater in Blender From Phone Photos</h3><p class="paragraph" style="text-align:left;">In Denoised, we watch Codex reconstruct a real theater in Blender from phone photos in about 10 minutes. It is a useful demonstration of how quickly a filmmaker can turn casual reference images into a workable starting point for planning shots, blocking, or testing an idea in 3D.</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/y_DIFhF3xXw" width="100%"></iframe><p class="paragraph" style="text-align:left;">The workflow is not a replacement for measured scans, clean production geometry, or an experienced artist’s judgment. But for early previs and location-driven exploration, it suggests a faster path from reference to scene.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://Nvidia.Read?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Read</a> the <a class="link" href="https://www.vp-land.com/p/ces-2026-fuji-s-fake-8mm-camera-lego-smart-bricks-atlas-robot-and-more?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">show notes</a> or watch the <a class="link" href="https://www.youtube.com/watch?v=y_DIFhF3xXw&utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">full episode</a>.</p><p class="paragraph" style="text-align:left;"><b>Watch/Listen & Subscribe</b></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://vpgo.link/denoised-spotify?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Spotify</a> | <a class="link" href="https://vpgo.link/denoised-apple?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Apple Podcast</a> | <a class="link" href="https://vpgo.link/denoised-youtube?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">YouTube</a></p></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1d8019c1-2e72-46e3-a35e-925e89e79f90/Clipboard_-_Box_Small_1_.png?t=1712401303"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#219c30;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">👔 Open Job Posts</h2><p class="paragraph" style="text-align:left;"><a class="link" href="https://job-boards.eu.greenhouse.io/creativefabrica/jobs/4856784101?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">AI Video Creator & Edito</a>r - Remote<br>Creative Fabrica</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://jobs.ashbyhq.com/runway-ml/5e8d6c9b-0d30-4bb2-90a9-df7d937aa44e?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Research Science Manager, Foundation Models</a> - Remote<br>Runway</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://job-boards.greenhouse.io/blackforestlabs?error=true&utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Forward Deployed Machine Learning Engineer</a> - San Francisco, CA<br>Black Forest Labs</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vfxengine.com/jobs/generative-ai-supervisor/907d1bc6-188d-48d7-a90a-87fe4b63b99c?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Generative AI Supervisor</a> - Milan, Italy<br>VFX Engine</p></div><hr class="content_break"><div id="events" class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/13731556-210f-41c3-af6d-0fe92e7099ae/Calendar_-_Box_Small.png?t=1712401007"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#3a8acc;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">📆 Upcoming Events</h2><p class="paragraph" style="text-align:left;"><b>Jul 19 to 23</b><br><a class="link" href="https://s2026.siggraph.org/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">SIGGRAPH 2026</a><br>Los Angeles, CA</p><p class="paragraph" style="text-align:left;"><b>Jul 20 to 22</b><br><a class="link" href="https://www.digitalhollywood.com?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">AI & Entertainment Experience / Super Creativity at Digital Hollywood</a><br>Virtual</p><p class="paragraph" style="text-align:left;"><b>July 24</b><br>🆕 <a class="link" href="https://www.vp-land.com/p/hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-se?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">HDR & Digital Color in Realtime Video Workflows: PIXERA’s Santa Monica Educational Summit Focuses on Seeing HDR Clearly</a><br>Santa Monica, CA</p><p class="paragraph" style="text-align:left;"><b>Aug 4 to 6</b><br><a class="link" href="https://ai4.io?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Ai4 2026</a><br>Las Vegas, NV</p><p class="paragraph" style="text-align:left;"><b>Aug 1 to Sep 5</b><br><a class="link" href="https://pages.becomecgpro.com/ai-for-filmmakers-course-live?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">AI for Filmmakers Live Course (CG Pro)</a><br>Virtual</p><p class="paragraph" style="text-align:left;"><b>Sept 11 to 14</b><br><a class="link" href="https://show.ibc.org?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">IBC 2026</a><br>Amsterdam, Netherlands</p><p class="paragraph" style="text-align:left;"><i>View the full event calendar and submit your own events </i><i><a class="link" href="https://www.vp-land.com/vp-events?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">here</a></i><i>. </i></p></div><hr class="content_break"><div id="section" class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1206dfb7-86ea-4af9-a638-7b591764cf0c/Martini_-_Box_Small.png?t=1738624339"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#b4d8db;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/JoseRMejia/status/2067241146623926525?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools"><p> Twitter tweet </p></a></blockquote></div><hr class="content_break"><hr class="content_break"><div class="section" style="background-color:transparent;border-color:#c40101;border-radius:50px;border-style:solid;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:20.0px 20.0px 20.0px 20.0px;"><p class="paragraph" style="text-align:left;"><b>Thanks for reading VP Land!</b></p><p class="paragraph" style="text-align:left;">Thanks for reading VP Land!</p><p class="paragraph" style="text-align:left;">Have a link to share or a story idea? <a class="link" href="https://newterritory.media/contact/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Send it here</a>.</p><p class="paragraph" style="text-align:left;">Interested in reaching media industry professionals? <a class="link" href="https://newterritory.media/advertise-with-vp-land/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-puts-ai-agents-inside-your-creative-tools" target="_blank" rel="noopener noreferrer nofollow">Advertise with us</a>.</p></div></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:subtitle>PLUS: Alibaba's 2.4T model, Google's free head rig, a Resolve agent</itunes:subtitle><itunes:author>Coffee and Celluloid</itunes:author><itunes:summary>PLUS: Alibaba's 2.4T model, Google's free head rig, a Resolve agent</itunes:summary><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>AI Agents Move Into Houdini, Unreal, and Adobe as NVIDIA Rallies Creative Apps Around MCP</title>
  <description></description>
      <enclosure length="326983" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/60938ab6-742d-4253-910e-a5dc53f6fb40/unnamed__24_.jpg"/>
  <link>https://newsletter.vp-land.com/p/ai-agents-move-into-houdini-unreal-and-adobe-as-nvidia-rallies-creative-apps-around-mcp</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/ai-agents-move-into-houdini-unreal-and-adobe-as-nvidia-rallies-creative-apps-around-mcp</guid>
  <pubDate>Mon, 20 Jul 2026 19:17:43 +0000</pubDate>
  <atom:published>2026-07-20T19:17:43Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Utility Ai]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">NVIDIA used its <a class="link" href="https://blogs.nvidia.com/blog/siggraph-news-2026/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=ai-agents-move-into-houdini-unreal-and-adobe-as-nvidia-rallies-creative-apps-around-mcp#mcp" target="_blank" rel="noopener noreferrer nofollow">SIGGRAPH announcements</a> to line up a wave of creative software behind Model Context Protocol (MCP), the open standard that lets an AI agent read an application&#39;s state and operate it directly. Tools from Adobe, Affinity, Blender, Boris FX, Foundry, SideFX, and Epic Games now expose MCP connections, putting agent control inside the apps artists already run.</p><ul><li><p class="paragraph" style="text-align:left;"><b>The shift is from chat to control.</b> Instead of generating an asset in a separate window, an agent can inspect a project, change settings, and render inside the host application.</p></li><li><p class="paragraph" style="text-align:left;"><b>Repetitive production work is the first target.</b> NVIDIA lists texture inspection, color-management validation, export-variant prep, playblasts, and shot validation as jobs an agent can take on.</p></li><li><p class="paragraph" style="text-align:left;"><b>NVIDIA supplies the plumbing, not the apps.</b> The company is positioning its Agent Toolkit, RTX PRO workstations, and DGX systems as the layer underneath.</p></li></ul><h2 class="heading" style="text-align:left;" id="mcp-lets-an-agent-read-a-project-an">MCP lets an agent read a project and call its functions, not guess from a screenshot</h2><p class="paragraph" style="text-align:left;">MCP gives an AI client a structured way to see an application&#39;s project state and call its functions, rather than working from pasted text or a captured image. We covered the pattern when <a class="link" href="https://www.vp-land.com/p/unreal-engine-5-8-embeds-an-mcp-server-so-ai-agents-can-drive-the-editor?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=ai-agents-move-into-houdini-unreal-and-adobe-as-nvidia-rallies-creative-apps-around-mcp" target="_blank" rel="noopener noreferrer nofollow">Unreal Engine embedded an MCP server</a> in its editor. NVIDIA&#39;s SIGGRAPH slate extends that same approach across a wider set of creative tools at once, standardizing on one protocol rather than a per-vendor plugin.</p><h2 class="heading" style="text-align:left;" id="seven-creative-apps-now-expose-mcp-">Seven creative apps now expose MCP connections</h2><ul><li><p class="paragraph" style="text-align:left;"><b>Adobe</b> is expanding its AI Assistant across Firefly, Express, and Creative Cloud, and shipping an Adobe Express Developer MCP Server so coding assistants can build Express add-ons against official documentation.</p></li><li><p class="paragraph" style="text-align:left;"><b>Affinity by Canva</b> added an AI Connector for Claude that uses MCP for natural-language automation, covering layer renaming, asset resizing, bulk edits, vector optimization, and file preparation.</p></li><li><p class="paragraph" style="text-align:left;"><b>Blender</b> offers a lightweight MCP server through Blender Lab, exposing a natural-language interface to Blender&#39;s Python API and documentation.</p></li><li><p class="paragraph" style="text-align:left;"><b>Boris FX Silhouette</b> ships an MCP server that lets an assistant inspect projects, build node trees, edit shapes and keyframes, and render frames.</p></li><li><p class="paragraph" style="text-align:left;"><b>Foundry Griptape</b> natively supports MCP, providing AI orchestration built for professional VFX pipelines; we <a class="link" href="https://www.vp-land.com/p/foundry-brings-griptape-ai-agents-into-nuke-blender-and-maya-via-mcp?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=ai-agents-move-into-houdini-unreal-and-adobe-as-nvidia-rallies-creative-apps-around-mcp" target="_blank" rel="noopener noreferrer nofollow">covered its Griptape agents</a> landing in Nuke and Blender.</p></li><li><p class="paragraph" style="text-align:left;"><b>SideFX Houdini</b> brings MCP support to Houdini 22 through its new APEX Script workflow, aimed at procedural character-rig generation.</p></li><li><p class="paragraph" style="text-align:left;"><b>Unreal Engine</b> connects AI clients to the Unreal Editor through MCP for workflows that interact with editor capabilities.</p></li></ul><h2 class="heading" style="text-align:left;" id="nvidia-supplies-the-agent-framework">NVIDIA supplies the agent framework and the hardware, not the creative apps</h2><p class="paragraph" style="text-align:left;">NVIDIA&#39;s role sits underneath the software. Its Agent Toolkit adds MCP client and server support so developers can wire agents to these tools, and the company points creators toward RTX PRO workstations plus DGX Spark and DGX Station systems for running models locally. Its Omniverse libraries are also exposed as agent-accessible tools. Virtual-production experiments already pointed at where this goes: <a class="link" href="https://www.vp-land.com/p/aximmetry-s-mcp-server-lets-ai-agents-build-virtual-production-scenes?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=ai-agents-move-into-houdini-unreal-and-adobe-as-nvidia-rallies-creative-apps-around-mcp" target="_blank" rel="noopener noreferrer nofollow">Aximmetry&#39;s MCP server</a> lets an agent build and adjust scenes from a plain-language request.</p><h2 class="heading" style="text-align:left;" id="a-shared-protocol-means-one-agent-c">A shared protocol means one agent can move from Houdini to Nuke to Unreal</h2><p class="paragraph" style="text-align:left;">The connective thread across these announcements is standardization. When Adobe, Foundry, SideFX, and Epic all expose the same protocol, an agent built once can move between applications instead of being locked to a single vendor&#39;s integration. For artists, the near-term payoff is offloading setup and cleanup work; the larger one is a toolchain where the same assistant can follow a project from Houdini to Nuke to Unreal. Availability varies by vendor and much of it is early, but the direction is now shared across the biggest names in production software.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>NVIDIA’s DGX Station Is the Home for Its Local AI Agent Stack</title>
  <description></description>
      <enclosure length="2568652" type="image/png" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/44e1698d-fa2c-46d4-b5fd-6ab2cbe6db4a/unnamed__1_.png"/>
  <link>https://newsletter.vp-land.com/p/nvidia-s-dgx-station-is-the-home-for-its-local-ai-agent-stack</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/nvidia-s-dgx-station-is-the-home-for-its-local-ai-agent-stack</guid>
  <pubDate>Mon, 20 Jul 2026 19:12:13 +0000</pubDate>
  <atom:published>2026-07-20T19:12:13Z</atom:published>
    <category><![CDATA[Hardware]]></category>
    <category><![CDATA[Article]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">NVIDIA is positioning DGX Station as the local hardware base for an open agent stack that can run large models, govern their behavior, and call simulation tools inside Blender. Announced at <a class="link" href="https://blogs.nvidia.com/blog/siggraph-news-2026/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-dgx-station-is-the-home-for-its-local-ai-agent-stack#nemoclaw-dgx-station" target="_blank" rel="noopener noreferrer nofollow">SIGGRAPH</a>, the stack combines NVIDIA’s Agent Toolkit with NemoClaw blueprints, Nemotron 3 Ultra, OpenShell, and Omniverse libraries.</p><p class="paragraph" style="text-align:left;">For technical teams that need to keep assets and workloads on-premises, the point is not simply a larger local model. NVIDIA is packaging the model, runtime, and tool connections so an agent can work against controlled, callable software functions rather than a cloud API.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Local agent stack.</b> DGX Station is designed to run Nemotron 3 Ultra, a 550-billion-parameter open model, on a single desktop system.</p></li><li><p class="paragraph" style="text-align:left;"><b>Governed execution.</b> OpenShell provides a sandboxed runtime with defined policies for agent behavior.</p></li><li><p class="paragraph" style="text-align:left;"><b>Callable simulation tools.</b> A NemoClaw blueprint connects agents in Blender to Omniverse libraries for RTX sensor simulation and physics work.</p></li></ul><h2 class="heading" style="text-align:left;" id="dgx-station-supplies-the-local-capa">DGX Station supplies the local capacity for NVIDIA’s agent stack</h2><p class="paragraph" style="text-align:left;">DGX Station runs on NVIDIA’s GB300 Grace Blackwell Ultra Desktop Superchip. NVIDIA says the system delivers up to 20 petaflops of FP4 compute and 748GB of coherent memory, capacity intended to make a 550-billion-parameter model practical on one desktop machine.</p><p class="paragraph" style="text-align:left;">That matters because the Agent Toolkit stack is not a single application. It combines the Nemotron 3 Ultra model with OpenShell, NemoClaw blueprints, and Omniverse libraries. NVIDIA describes each piece as open:</p><ul><li><p class="paragraph" style="text-align:left;"><b>Nemotron 3 Ultra</b> is the 550-billion-parameter model tuned for DGX Station hardware.</p></li><li><p class="paragraph" style="text-align:left;"><b>OpenShell</b> is an open source runtime intended to sandbox agents and apply defined policies.</p></li><li><p class="paragraph" style="text-align:left;"><b>NemoClaw</b> provides open blueprints for assembling custom autonomous agents from the model, harness, and runtime.</p></li><li><p class="paragraph" style="text-align:left;"><b>Omniverse libraries</b> expose physics simulation and 3D asset workflow functions as tools an agent can call.</p></li></ul><p class="paragraph" style="text-align:left;">The practical proposition is a stack that technical teams can inspect, adapt, and keep within their own infrastructure.</p><p class="paragraph" style="text-align:left;">NVIDIA says a ConnectX-8 SuperNIC provides up to 800GB/s of bandwidth and can link two DGX Stations. The company has published playbooks for using two systems to run larger models or support more concurrent users. It also says local operation can be set up in three steps in about 30 minutes, without an internet connection.</p><p class="paragraph" style="text-align:left;">DGX Station is available to order from ASUS, Dell Technologies, Exxact, GIGABYTE, HP, MSI, and Supermicro. NVIDIA did not announce pricing.</p><h2 class="heading" style="text-align:left;" id="nemo-claw-makes-blenders-simulation">NemoClaw makes Blender’s simulation tools callable by an agent</h2><p class="paragraph" style="text-align:left;">The creative-software angle arrives through a NemoClaw blueprint that integrates Omniverse libraries into Blender. It gives an agent callable RTX sensor-simulation and physics tools within an artist’s existing scene, allowing the agent to initiate simulation work rather than merely suggest steps in a chat window.</p><p class="paragraph" style="text-align:left;">The setup follows a broader shift toward AI agents operating applications through explicit tool connections. We covered <a class="link" href="https://www.vp-land.com/p/foundry-brings-griptape-ai-agents-into-nuke-blender-and-maya-via-mcp?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-dgx-station-is-the-home-for-its-local-ai-agent-stack" target="_blank" rel="noopener noreferrer nofollow">Foundry’s Griptape integration</a>, which brought agents into Nuke, Blender, and Maya through the Model Context Protocol.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/aximmetry-s-mcp-server-lets-ai-agents-build-virtual-production-scenes?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-dgx-station-is-the-home-for-its-local-ai-agent-stack" target="_blank" rel="noopener noreferrer nofollow">Aximmetry’s MCP server</a> offered another version of the idea, letting agents build and adjust virtual-production scenes from a plain-language request. NVIDIA’s approach ties that application-level access to its own local model, runtime, and DGX Station hardware.</p><h2 class="heading" style="text-align:left;" id="the-studio-question-is-governance-n">The studio question is governance, not only speed</h2><p class="paragraph" style="text-align:left;">For studios working with unreleased assets, scripts, or client material, local execution addresses a basic operational constraint: whether creative data has to leave the building for an agent to be useful. A DGX Station running an offline model, with OpenShell enforcing policies and Blender exposing limited callable tools, gives teams more control over where that work happens and what the agent is allowed to do.</p><p class="paragraph" style="text-align:left;">The substantive test is whether the tools remain reliable on production scenes. NVIDIA’s published specifications establish the hardware capacity and the software components, but pricing and performance with real studio workloads remain open questions.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>We Watched Codex Rebuild a Real Theater in Blender From Phone Photos</title>
  <description></description>
      <enclosure length="269060" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/de279fa0-f576-4d1c-aeb1-b58432cf21ac/unnamed__23_.jpg"/>
  <link>https://newsletter.vp-land.com/p/we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos</guid>
  <pubDate>Mon, 20 Jul 2026 17:52:53 +0000</pubDate>
  <atom:published>2026-07-20T17:52:53Z</atom:published>
    <category><![CDATA[Denoised Podcast]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">This episode of Denoised starts with a simple test that lands somewhere bigger: feeding a handful of cell-phone photos to OpenAI&#39;s Codex and asking it to reconstruct a real building in Blender. From there we get into the AI Odyssey trailer everyone argued about, Foundry&#39;s SmartRoto for Nuke, Seedance 2.5&#39;s pricey teaser, and America&#39;s first genuinely competitive open-weights model. It also serves as our SIGGRAPH primer, since the show is landing in LA and we will both be there.</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/y_DIFhF3xXw" width="100%"></iframe><div class="section" style="background-color:transparent;margin:30.0px 30.0px 30.0px 30.0px;padding:0.0px 0.0px 0.0px 0.0px;"><table width="100%" class="bh__column_wrapper"><tr><td width="33%" class="bh__column"><div class="image"><a class="image__link" href="https://vpgo.link/denoised-spotify?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/73810210-52ee-4c5a-bfc9-ecfbf1e97f7a/Spotify_-_04.png?t=1739592377"/></a></div></td><td width="33%" class="bh__column"><div class="image"><a class="image__link" href="https://vpgo.link/denoised-apple?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5cf3881d-77c4-4364-a1d6-679d33cf5734/Apple_-_04.png?t=1739592405"/></a></div></td><td width="33%" class="bh__column"><div class="image"><a class="image__link" href="https://vpgo.link/denoised-youtube?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ce4801d4-82ec-413c-bdb9-24813d091e59/YouTube_-_04.png?t=1739592417"/></a></div></td></tr></table></div><h2 class="heading" style="text-align:left;" id="quick-take">Quick Take</h2><p class="paragraph" style="text-align:left;">One quick test set the theme for the whole episode: a coding agent took cell-phone snapshots of a theater and rebuilt it as editable Blender geometry in roughly 10 minutes. That reframes a bigger argument we keep having on the show about where AI fits into a 3D pipeline. Is the future fully synthetic generation, or is it 3D doing the precise structural work while AI handles the final look? The AI Odyssey trailer, Seedance 2.5, and a new wave of open-weights models all push on the same question from different angles.</p><h2 class="heading" style="text-align:left;" id="what-were-watching-siggraph-lands-i">What We&#39;re Watching: SIGGRAPH Lands in LA</h2><p class="paragraph" style="text-align:left;">SIGGRAPH runs in LA, which makes it easy to reach for people in media and entertainment. Addy&#39;s framing: it is NAB for the research crowd, where university labs and companies show work that can be five or ten years out. Netflix already teased some of the papers it plans to present, and the density of AI researchers makes the show a recruiting event as much as a technical one.</p><p class="paragraph" style="text-align:left;">Two satellite events are worth flagging. The <a class="link" href="https://www.vp-land.com/p/dreamworks-moonray-keynote-headlines-aswf-open-source-days-in-la?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">ASWF Open Source Days</a> sessions run on Sunday and dig into the open standards, like USD and OpenTimelineIO, that this episode keeps circling back to.</p><p class="paragraph" style="text-align:left;">There is also a free <a class="link" href="https://www.eventbrite.com/e/ai-workflows-summit-tickets-1993363191967?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">AI Workflows Summit</a> on Tuesday evening that does not require a SIGGRAPH pass, with talks on static Gaussian splats from NVIDIA, real-time volumetric workflows, and a ComfyUI agentic tool called <a class="link" href="https://ComfyCode.ai?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">ComfyCode.ai</a> that runs locally and hooks into production-grade pipeline standards. We are moderating the closing panel there.</p><h2 class="heading" style="text-align:left;" id="what-we-tested-codex-rebuilt-the-cu">What We Tested: Codex Rebuilt the Culver Theater in Blender</h2><p class="paragraph" style="text-align:left;">Here is the test that opened the episode. We pointed OpenAI&#39;s <a class="link" href="https://openai.com/codex?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Codex</a>, its coding agent and answer to Claude Code, at Blender and asked it to install the tooling itself.</p><p class="paragraph" style="text-align:left;">The agent found the third-party <a class="link" href="https://www.blender.org/lab/mcp-server?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Blender MCP</a> server, set it up, and connected to Blender without hand-holding.</p><p class="paragraph" style="text-align:left;">Then came the real ask: rebuild the Culver Theater and its surrounding block using a few phone photos shot at AI on the Lot. These were not ideal reference images, mostly one-sided, with no clean angle on the corner. The Art Deco detailing (curved forms, the sphere held up on top, intricate trim) is exactly the geometry a human modeler would dread.</p><p class="paragraph" style="text-align:left;">The result, after one revision pass asking for a more rounded corner door, was a strong starting point rather than a finished asset.</p><ul><li><p class="paragraph" style="text-align:left;"><b>What it got:</b> the overall massing, the lamps, the rounded corners, and a believable interpretation of the Art Deco floor stamping and side-wall movie posters that were barely visible in the source photos.</p></li><li><p class="paragraph" style="text-align:left;"><b>What it missed:</b> the theater marquee sign was the most obvious error, and it skipped several of the door windows.</p></li><li><p class="paragraph" style="text-align:left;"><b>Why it still matters:</b> the geometry came in cleanly separated in the outliner, ready to modify. It even generated a camera move and, on request, relit the scene as a moody night exterior, though some emissive values landed on trim that should not glow.</p></li></ul><p class="paragraph" style="text-align:left;">The whole build took around 10 minutes once the server spun up, and the agent felt notably more responsive driving the computer than screen-reading control loops we have used before.</p><h2 class="heading" style="text-align:left;" id="where-3-d-becomes-the-pilot-and-ai-">Where 3D Becomes the Pilot and AI Becomes the Renderer</h2><p class="paragraph" style="text-align:left;">The test crystallized a prediction we have been building toward. For our own work, photoreal accuracy is not the point; a gray-box model that gets you 70% of the way there gives you real geometry to run camera moves and feed into video-to-video passes, with generative cleanup handling the rest.</p><p class="paragraph" style="text-align:left;">Addy&#39;s larger read: a year or two ago the assumption was that generative AI would pull production from 3D back to flat 2D, because why build in 3D when you can reach final pixel directly. That was wrong. You still need 3D for control, for the specific camera move, for the sphere held up by four supports. The workflow taking shape is 3D as the pilot doing the precise structural work, and AI as the renderer filling the photorealism gap. As GPUs push from billions of polygons toward trillions and inference keeps getting faster, the pitch is an AI 3D operator that handles modeling, shading, and rigging under the hood while a director just talks to the scene.</p><h2 class="heading" style="text-align:left;" id="what-we-explored-a-new-video-format">What We Explored: A New Video Format Built From a Prompt</h2><p class="paragraph" style="text-align:left;">If an agent can install its own tools and script complex software, it can also invent new plumbing. Alex Barashkov used Codex to build <a class="link" href="https://www.vp-land.com/p/aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Aval</a>, an open-source format for state-driven interactive video on the web, with small file sizes, low CPU overhead, alpha transparency, and a web-native runtime.</p><p class="paragraph" style="text-align:left;">That is the part worth sitting with. The barrier to inventing a new file format or delivery standard used to be enormous; USD took Pixar the better part of a decade. The risk is a Wild West where every studio vibe-codes its own incompatible pipeline, which is why open standards matter more, not less, as building custom tooling gets cheap.</p><h2 class="heading" style="text-align:left;" id="what-we-debated-smart-roto-brings-a">What We Debated: SmartRoto Brings AI Into a Pro Roto Pipeline</h2><p class="paragraph" style="text-align:left;">Foundry announced <a class="link" href="https://www.vp-land.com/p/foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">SmartRoto</a> for Nuke, an AI-assisted rotoscoping add-on the company says can be up to 4x faster while keeping artists in control. Similar masking and tracking already lives in Resolve and After Effects, but Foundry has historically been careful about how it frames AI for high-end professionals.</p><p class="paragraph" style="text-align:left;">That care is notable given Foundry now owns Griptape and is leaning into agentic pipelines. Griptape&#39;s studio push, which we covered when Foundry brought its <a class="link" href="https://www.vp-land.com/p/foundry-brings-griptape-ai-agents-into-nuke-blender-and-maya-via-mcp?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">agents into Nuke, Blender, and Maya</a>, emphasizes OCIO color support and local GPU inference. That local angle addresses a real studio anxiety about pushing unreleased IP to the cloud.</p><p class="paragraph" style="text-align:left;">We landed on a live question worth testing: AI generations carry inherent noise, and the open debate is whether that hurts rotoscoping or, more likely, camera tracking, which has to solve a 2D image back into a 3D camera move.</p><h2 class="heading" style="text-align:left;" id="what-we-questioned-does-an-ai-gener">What We Questioned: Does an AI-Generated Odyssey Prove Anything?</h2><p class="paragraph" style="text-align:left;">An <a class="link" href="https://www.hollywoodreporter.com/business/digital/ai-generated-feature-odysseus-the-fall-1236646751?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">AI-generated feature trailer</a> built for a few thousand dollars got framed as going head-to-head with Christopher Nolan&#39;s $250 million adaptation. The reality check: the clip runs under two minutes, so it is a proof of concept, not a competing feature.</p><p class="paragraph" style="text-align:left;">Some individual shots look genuinely good, well color-graded and striking on their own. The breakdowns show up in the connective tissue. AI still cannot nail real lensing and glass, so perspective and distance shots fall apart. Camera geography and placement drift, and fully synthetic performances stay in the uncanny valley.</p><p class="paragraph" style="text-align:left;">Addy&#39;s recurring argument is the useful takeaway: the real unlock is hybrid, combining AI with physical performance and a real cinema camera and lens for the reference plate, then letting AI fill in the world around it. Solve for the expensive parts of production and cut costs in half, and that is the win. Betting on fully synthetic everything trades away quality that the technology cannot yet recover.</p><h2 class="heading" style="text-align:left;" id="what-we-watched-seedance-25-s-tease">What We Watched: Seedance 2.5&#39;s Teaser and the $5-a-Shot Problem</h2><p class="paragraph" style="text-align:left;">On the generation side, ByteDance released a teaser made with <a class="link" href="https://ai.byteplus.com/lumina/en/resource/what-is-seedance-2-5?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Seedance 2.5</a>, a soccer-themed spot following a kid dribbling a ball through London. The ball-to-foot contact, normally a hard problem for AI, holds up well enough that we may end up eating our earlier skepticism, which is the whole point of tracking this stuff in the open.</p><p class="paragraph" style="text-align:left;">Two caveats stand out. The model reportedly supports up to 30 reference images, which raises an open directing problem: whether it can place all those references correctly in space from a wide shot plus close-ups. And the cost is steep. A maxed-out 4K shot of a few seconds runs roughly $4 to $5, which makes real iteration expensive. At those prices, shooting something practically can be cheaper than blasting a model at it, a tradeoff that already pushed some software firms to rehire humans over pricey tokens.</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="what-we-explored-america-finally-ha">What We Explored: America Finally Has a Competitive Open-Weights Model</h2><p class="paragraph" style="text-align:left;">In the open-source lane, Mira Murati&#39;s Thinking Machines released its first model, <a class="link" href="https://www.vp-land.com/p/thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Inkling</a>, an open-weights, multimodal model built to be fine-tuned. It is not at the level of the top closed frontier models, but it is a meaningful leap as an American open model in a category that has been dominated by Chinese labs like DeepSeek.</p><p class="paragraph" style="text-align:left;">A few standout traits from the discussion:</p><ul><li><p class="paragraph" style="text-align:left;"><b>Built to be fine-tuned.</b> Addy flagged a feature he had not seen elsewhere: the model can fine-tune itself on your data, so an agent running a specialized use case can adapt without the current markdown-and-context-window workaround.</p></li><li><p class="paragraph" style="text-align:left;"><b>A large context window</b> and a free playground to test it, with third-party providers likely to add hosting support.</p></li><li><p class="paragraph" style="text-align:left;"><b>A momentum play.</b> Open-weighting the model is a way for a newcomer to build a user flywheel rather than trying to out-scale ChatGPT or Claude head-on, much like early Stable Diffusion did with image models.</p></li></ul><p class="paragraph" style="text-align:left;">That ties into a bigger watch list: Ilya Sutskever&#39;s and Yann LeCun&#39;s new ventures. LeCun&#39;s bet, per Addy, is that LLM scaling is hitting diminishing returns and the JEPA architecture points toward a more fluid model where text, image, and video inputs blend rather than getting stitched together. Meanwhile, Kimi 3 is teased at a reported 2 to 3 trillion parameters as an open-weights release, echoing the scale of Alibaba&#39;s <a class="link" href="https://www.vp-land.com/p/alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">2.4-trillion-parameter Qwen preview</a>. The uncomfortable note we ended on: the biggest open model may again ship from China rather than the US.</p><h2 class="heading" style="text-align:left;" id="bottom-line-3-d-structure-ai-finish">Bottom Line: 3D Structure, AI Finish</h2><p class="paragraph" style="text-align:left;">The through-line across every story is the same split between structure and finish.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Codex plus Blender MCP</b> shows AI can now handle the tedious software work of building and lighting a 3D scene, turning phone photos into editable geometry that feeds a hybrid pipeline.</p></li><li><p class="paragraph" style="text-align:left;"><b>The AI Odyssey and Seedance 2.5</b> show pure generation nails individual shots but still breaks on lensing, continuity, and cost, which is why real cameras and physical performance stay in the loop.</p></li><li><p class="paragraph" style="text-align:left;"><b>Inkling and the open-weights wave</b> show the models underneath are getting cheaper to adapt and self-host, with fine-tuning and local inference mattering more as API costs climb.</p></li></ul><p class="paragraph" style="text-align:left;">The pattern to watch is not whether AI replaces 3D or live action, but how quickly the agent layer makes both easier to control.</p><h2 class="heading" style="text-align:left;" id="links-from-this-episode">Links from This Episode</h2><p class="paragraph" style="text-align:left;"><b>Tools & Platforms:</b></p><ul><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://openai.com/codex?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">OpenAI Codex</a>, the coding agent used to run the Blender test</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.blender.org/lab/mcp-server?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Blender MCP</a>, the server that let Codex drive Blender</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://ai.byteplus.com/lumina/en/resource/what-is-seedance-2-5?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Seedance 2.5</a>, ByteDance&#39;s teased video model</p></li></ul><p class="paragraph" style="text-align:left;"><b>Tools & Releases:</b></p><ul><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Foundry SmartRoto for Nuke</a>, VP Land coverage</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/foundry-brings-griptape-ai-agents-into-nuke-blender-and-maya-via-mcp?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Foundry Griptape agents in Nuke, Blender, and Maya</a>, VP Land coverage</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Thinking Machines Inkling</a>, VP Land coverage</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Aval interactive video format</a>, VP Land coverage</p></li></ul><p class="paragraph" style="text-align:left;"><b>News & Analysis:</b></p><ul><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.hollywoodreporter.com/business/digital/ai-generated-feature-odysseus-the-fall-1236646751?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">AI-generated Odyssey trailer</a>, The Hollywood Reporter</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">Alibaba&#39;s 2.4-trillion-parameter Qwen preview</a>, VP Land coverage</p></li></ul><p class="paragraph" style="text-align:left;"><b>Events:</b></p><ul><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://s2026.siggraph.org?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">SIGGRAPH</a>, official site</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/dreamworks-moonray-keynote-headlines-aswf-open-source-days-in-la?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">ASWF Open Source Days</a>, VP Land coverage</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.eventbrite.com/e/ai-workflows-summit-tickets-1993363191967?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">AI Workflows Summit</a>, free SIGGRAPH satellite event</p></li></ul><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.youtube.com/watch?v=y_DIFhF3xXw&utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-watched-codex-rebuild-a-real-theater-in-blender-from-phone-photos" target="_blank" rel="noopener noreferrer nofollow">WATCH THE FULL EPISODE ON YOUTUBE →</a></b></p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>NVIDIA's Synthetic Video Detector Puts a Deepfake Score Inside Newsroom Workflows</title>
  <description></description>
      <enclosure length="292433" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/727a7151-069b-42bd-8777-8b743288fdbd/unnamed__22_.jpg"/>
  <link>https://newsletter.vp-land.com/p/nvidia-s-synthetic-video-detector-puts-a-deepfake-score-inside-newsroom-workflows</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/nvidia-s-synthetic-video-detector-puts-a-deepfake-score-inside-newsroom-workflows</guid>
  <pubDate>Mon, 20 Jul 2026 17:48:18 +0000</pubDate>
  <atom:published>2026-07-20T17:48:18Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Utility Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">NVIDIA introduced the Synthetic Video Detector NIM microservice at SIGGRAPH, adding an AI-assisted signal that flags whether a video clip contains synthetic content. It is part of the <a class="link" href="https://blogs.nvidia.com/blog/siggraph-news-2026/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-synthetic-video-detector-puts-a-deepfake-score-inside-newsroom-workflows#synthetic-video" target="_blank" rel="noopener noreferrer nofollow">NVIDIA AI for Media</a> platform and is aimed at editorial and media teams making fast calls on questionable footage.</p><ul><li><p class="paragraph" style="text-align:left;"><b>The tool produces a classifier score, not a verdict.</b> It analyzes video frame by frame and outputs a probability that the clip is synthetic, which teams can use to prioritize, flag, quarantine, or escalate.</p></li><li><p class="paragraph" style="text-align:left;"><b>Accuracy reached up to 92% on uncompressed video</b> in NVIDIA testing, holding at 82% even after heavy compression.</p></li><li><p class="paragraph" style="text-align:left;"><b>Wowza is already embedding it</b> for real-time detection across more than 35,000 livestreaming deployments.</p></li></ul><h2 class="heading" style="text-align:left;" id="a-framebyframe-score-built-to-survi">A frame-by-frame score built to survive newsroom compression</h2><p class="paragraph" style="text-align:left;">The microservice analyzes video frame by frame to generate a classifier score for synthetic content. Editorial teams can use that score to prioritize clips for review, flag or quarantine questionable footage, or escalate it for deeper analysis. NVIDIA positions the tool as an additional signal for time-sensitive decisions rather than a replacement for established verification practices.</p><p class="paragraph" style="text-align:left;">Detection quality usually degrades once a clip is compressed, resized, cropped, or re-encoded, which is exactly what happens to video moving through newsroom and social pipelines. NVIDIA says the model stays effective through those steps. In its testing, accuracy reached up to 92% on uncompressed video, 87% at 15% compression, and 82% at 50% compression.</p><p class="paragraph" style="text-align:left;">The latest model revision reported an AUC of 0.9614 and accuracy of 0.9453 on NVIDIA&#39;s internal test set. AUC, or Area Under the Curve, measures how well a classifier ranks positive samples above negative ones independent of thresholds. Those thresholds can be tuned, including more conservative settings that reduce the chance a synthetic clip slips through unflagged.</p><h2 class="heading" style="text-align:left;" id="built-for-realtime-review-at-newsro">Built for real-time review at newsroom speed</h2><p class="paragraph" style="text-align:left;">Speed matters when a clip needs a decision before it airs or posts. The microservice can process 1080p video in as little as 22 milliseconds on NVIDIA RTX systems and roughly 30 milliseconds on NVIDIA L40 GPUs.</p><p class="paragraph" style="text-align:left;">The detector ships in NVIDIA&#39;s <a class="link" href="https://www.vp-land.com/p/nvidia-nim-microservices?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-synthetic-video-detector-puts-a-deepfake-score-inside-newsroom-workflows" target="_blank" rel="noopener noreferrer nofollow">NIM microservice format</a>, the packaged deployment model we covered when NVIDIA first brought it to production media pipelines. That packaging is what lets editorial teams drop the model into an existing workflow instead of standing up a separate detection stack.</p><h2 class="heading" style="text-align:left;" id="onprem-and-airgapped-deployment-for">On-prem and air-gapped deployment for sensitive footage</h2><p class="paragraph" style="text-align:left;">Organizations can run the microservice closer to where sensitive video is captured, stored, or distributed, including on-premises, edge, hybrid, and approved air-gapped environments. NVIDIA frames that flexibility as a way for teams to keep control over video data, access, and operations.</p><p class="paragraph" style="text-align:left;">That deployment model targets a specific set of buyers. NVIDIA names broadcasters, government agencies, financial institutions, and critical infrastructure operators as the groups most exposed to synthetic media risk, and also the ones facing strict requirements around data residency, security, and operational control.</p><p class="paragraph" style="text-align:left;">Partner adoption is where the detector moves from a model to deployable infrastructure. Wowza is embedding the microservice through its Video Intelligence Framework, which we saw the company <a class="link" href="https://www.vp-land.com/p/lucidlink-s-connect-streams-files-from-frame-io-google-drive-and-dropbox-into-a-single-unified-files?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-synthetic-video-detector-puts-a-deepfake-score-inside-newsroom-workflows" target="_blank" rel="noopener noreferrer nofollow">introduce at NAB</a>, bringing real-time synthetic video detection into livestreaming workflows that span more than 35,000 deployments across over 170 countries. Pairing the detector with a video layer customers already run puts AI-assisted verification closer to ingest and streaming operations while keeping footage inside their own environments.</p><h2 class="heading" style="text-align:left;" id="detection-tooling-catches-up-to-the">Detection tooling catches up to the generation curve</h2><p class="paragraph" style="text-align:left;">The release lands as synthetic video quality keeps climbing and detection shifts from a research topic toward a working newsroom requirement. We covered how quickly <a class="link" href="https://www.vp-land.com/p/we-tested-switchx-live-while-seedance-sparked-a-deepfake-crisis?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-synthetic-video-detector-puts-a-deepfake-score-inside-newsroom-workflows" target="_blank" rel="noopener noreferrer nofollow">a single AI clip can trigger a public trust crisis</a>, the kind of scenario a frame-level score wired into ingest is meant to catch earlier. NVIDIA is positioning the Synthetic Video Detector as one signal inside a broader verification process, which sets a realistic bar for what automated detection can do as generation tools continue to improve.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>The ASC Is Open-Sourcing a Virtual Production Test Kit for LED Volumes</title>
  <description></description>
      <enclosure length="311068" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/16cfbe11-8091-4725-960a-0ec592e2acf1/unnamed__21_.jpg"/>
  <link>https://newsletter.vp-land.com/p/the-asc-is-open-sourcing-a-virtual-production-test-kit-for-led-volumes</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/the-asc-is-open-sourcing-a-virtual-production-test-kit-for-led-volumes</guid>
  <pubDate>Sun, 19 Jul 2026 22:03:08 +0000</pubDate>
  <atom:published>2026-07-19T22:03:08Z</atom:published>
    <category><![CDATA[Vfx Industry]]></category>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Virtual Production]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">The American Society of Cinematographers is releasing <b>StEM3-VP</b>, a set of open reference assets built to evaluate and calibrate in-camera visual effects stages, and the material is joining the Academy Software Foundation&#39;s <a class="link" href="https://dpel.aswf.io/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=the-asc-is-open-sourcing-a-virtual-production-test-kit-for-led-volumes" target="_blank" rel="noopener noreferrer nofollow">Digital Production Example Library</a>.</p><ul><li><p class="paragraph" style="text-align:left;"><b>StEM3-VP targets LED volumes.</b> The assets cover virtual production, in-camera visual effects, and large LED walls, the parts of a modern set that are hardest to prove out before a shoot.</p></li><li><p class="paragraph" style="text-align:left;"><b>It is openly available.</b> The project lands in DPEL, the ASWF library that hosts production-grade sample content for testing hardware and software.</p></li><li><p class="paragraph" style="text-align:left;"><b>It continues a two-decade lineage.</b> StEM3-VP follows the ASC&#39;s original StEM from 2004 and StEM2 in 2022.</p></li></ul><p class="paragraph" style="text-align:left;">The work comes from the ASC&#39;s Motion Imaging Technology Council, the same technical body behind the society&#39;s earlier evaluation films.</p><h2 class="heading" style="text-align:left;" id="what-st-em-3-vp-is-built-to-do-on-a">What StEM3-VP is built to do on an ICVFX stage</h2><p class="paragraph" style="text-align:left;">StEM3-VP is Standard Evaluation Material aimed squarely at LED-based in-camera visual effects. According to the ASC Motion Imaging Technology Council, the goal is to give productions a suite of tools to evaluate and verify LED ICVFX stages before principal photography begins.</p><p class="paragraph" style="text-align:left;">The material is meant to aid the processing, evaluation, and calibration of the image path between the LED wall and the camera. That is the exact seam where color, moiré, and brightness problems surface on a volume, and where a bad match between wall output and camera capture can quietly cost a production time on the day.</p><p class="paragraph" style="text-align:left;">The working group frames the assets as a way to standardize the setup and proving of LED walls for a given production, using openly available media and tools rather than proprietary test patterns that vary from vendor to vendor.</p><h2 class="heading" style="text-align:left;" id="why-an-open-shared-reference-matter">Why an open, shared reference matters for volumes</h2><p class="paragraph" style="text-align:left;">LED volumes are expensive to book and unforgiving to misconfigure. A common, openly licensed reference gives cinematographers, DITs, and stage engineers a fixed target to test against, so a stage can be characterized and signed off before the crew arrives.</p><p class="paragraph" style="text-align:left;">Because the assets live in DPEL, they sit alongside other production-grade test content the Academy Software Foundation maintains for hardware and software development. We covered the <a class="link" href="https://www.vp-land.com/p/dreamworks-moonray-keynote-headlines-aswf-open-source-days-in-la?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=the-asc-is-open-sourcing-a-virtual-production-test-kit-for-led-volumes" target="_blank" rel="noopener noreferrer nofollow">Open Source Days program</a> where StEM3-VP was slated to be presented, part of a schedule spanning color fidelity standards, OpenPBR, and studio-safe AI workflows.</p><p class="paragraph" style="text-align:left;">The lineage gives working professionals a known quantity. The original StEM helped calibrate digital cinema projection in 2004. StEM2, released in 2022 under the ASWF Digital Assets License, shipped as a short film in QuickTime, DCP, IMF, and EXR formats, with SDR Rec709 and HDR Rec2020 PQ versions running from 48 to 1000 nits, specifically to stress a color pipeline. StEM3-VP extends that same evaluation approach to the LED stage.</p><h2 class="heading" style="text-align:left;" id="the-cinematographers-and-technologi">The cinematographers and technologists behind it</h2><p class="paragraph" style="text-align:left;">StEM3-VP is led by a working group of ASC members, color scientists, technologists, and manufacturers. Named contributors include <b>David Morin</b>, who co-chairs the effort, alongside Michael Goi, ASC, Jay Holben, Curtis Clark, ASC, Wendy Aylsworth, Joachim Zell, Rod Bogart, Gary Mandle, and Tim Kang.</p><p class="paragraph" style="text-align:left;">The DPEL integration was presented at the Academy Software Foundation&#39;s Open Source Days by Jim Geduldick of Spaceboy Labs and Jay Holben, an independent director and producer.</p><p class="paragraph" style="text-align:left;">Curtis Clark, ASC, a named contributor to StEM3-VP, has long led the society&#39;s motion-imaging standards work. For readers tracking how the ASC has been widening access to its technical resources, this follows the society&#39;s push into education with <a class="link" href="https://www.vp-land.com/p/asc-launches-asc-industry-s-most-prestigious-cinematography-learning-platform?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=the-asc-is-open-sourcing-a-virtual-production-test-kit-for-led-volumes" target="_blank" rel="noopener noreferrer nofollow">ASC+</a>, its subscription learning platform.</p><h2 class="heading" style="text-align:left;" id="what-this-means-for-stages-and-the-">What this means for stages and the crews that run them</h2><p class="paragraph" style="text-align:left;">For anyone commissioning or operating an LED volume, StEM3-VP offers a neutral yardstick that does not belong to a single wall vendor or processor manufacturer. That lowers the risk of walking onto a stage that looks fine in a demo and falls apart under a specific camera and color pipeline.</p><p class="paragraph" style="text-align:left;">The assets are openly available through DPEL, so stages can adopt them without licensing friction. As volumes spread beyond the largest studios, a shared calibration reference makes in-camera visual effects more predictable for the productions that can least afford a surprise on set.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>HDR &amp; Digital Color in Realtime Video Workflows: PIXERA’s Santa Monica Educational Summit Focuses on Seeing HDR Clearly</title>
  <description></description>
      <enclosure length="261694" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1f6ede80-a7e8-447b-a834-19f6a8f2141d/unnamed__20_.jpg"/>
  <link>https://newsletter.vp-land.com/p/hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-se</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-se</guid>
  <pubDate>Sun, 19 Jul 2026 21:59:30 +0000</pubDate>
  <atom:published>2026-07-19T21:59:30Z</atom:published>
    <category><![CDATA[Software]]></category>
    <category><![CDATA[Hardware]]></category>
    <category><![CDATA[Article]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">PIXERA is hosting a one-day educational summit on HDR and digital color at its Santa Monica studio on Friday, July 24. Titled <a class="link" href="https://www.eventbrite.com/e/hdr-digital-color-in-realtime-video-workflows-tickets-1988676802851?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-seeing-hdr-clearly" target="_blank" rel="noopener noreferrer nofollow">HDR & Digital Color in Realtime Video Workflows</a>, the program is built around talks, live demonstrations, and a reference system that lets attendees examine how HDR images are measured and perceived in real-time video workflows.</p><ul><li><p class="paragraph" style="text-align:left;"><b>PIXERA’s program runs from 9 a.m. to 5 p.m. PT, with up to 60 in-room attendees and a YouTube livestream.</b></p></li><li><p class="paragraph" style="text-align:left;"><b>The educational sessions use an HDR reference system that includes an LED wall, cinema cameras, vectorscopes, and PIXERA software.</b></p></li><li><p class="paragraph" style="text-align:left;"><b>The speaker program includes Paul Debevec, Tim S. Kang, Conor McGill, Camon Crocker, Angel Banchs, Damein Futch, Michael Kohler, and Kris Murray.</b></p></li></ul><h2 class="heading" style="text-align:left;" id="an-educational-day-about-what-hdr-m">An educational day about what HDR means in practice</h2><p class="paragraph" style="text-align:left;">The summit’s premise is straightforward: HDR is not only a delivery label or a number on a monitor specification. Its value depends on how an image is captured, displayed, measured, and ultimately seen. PIXERA frames the day around those questions, pairing presentations with live demonstrations so the discussion can stay connected to actual signal paths and viewing conditions.</p><p class="paragraph" style="text-align:left;">For virtual production and real-time teams, that makes the subject practical. HDR workflows can involve the LED display, camera, color pipeline, monitoring, and final delivery format at once. A mismatch in any part of that chain can change the image that a crew sees or the result that reaches an audience.</p><h2 class="heading" style="text-align:left;" id="a-reference-system-for-demonstratio">A reference system for demonstrations, not a new stage build</h2><p class="paragraph" style="text-align:left;">The Santa Monica event will use an HDR reference system comprising an LED wall, cinema cameras, vectorscopes, and PIXERA at the center. It is a teaching and demonstration environment for the summit, not an announcement of a newly built production stage.</p><p class="paragraph" style="text-align:left;">That distinction matters. The value of the setup is that presenters can discuss HDR with common visual and measurement references in the room. Vectorscopes can help make color information visible while the LED wall and cameras provide a real-time context for the conversation.</p><p class="paragraph" style="text-align:left;">PIXERA’s involvement also places the session near the company’s existing real-time media-server work. VP Land previously covered how <a class="link" href="https://www.vp-land.com/p/pixera-turns-every-led-wall-into-a-live-interactive-environment-with-one-platform?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-seeing-hdr-clearly" target="_blank" rel="noopener noreferrer nofollow">PIXERA turns an LED wall into a live interactive environment</a>, where the software is used to coordinate media playback and display systems. The July program narrows the focus to HDR and color education.</p><h2 class="heading" style="text-align:left;" id="a-program-shaped-by-imaging-and-col">A program shaped by imaging and color specialists</h2><p class="paragraph" style="text-align:left;">The agenda includes a session titled “Luminance in HDR: Relative or Absolute?” as well as a closing discussion, “HDR: Problems & Solutions in Creative Workflows.” Those titles point to the central tension in HDR work: technical measurements matter, but so do the viewing environment and the creative decisions that make an image readable.</p><p class="paragraph" style="text-align:left;">Paul Debevec is among the scheduled speakers. He is chief research officer at Netflix’s Eyeline Studios and is known for work in image-based lighting and computer graphics. The rest of the program brings together practitioners working across color, imaging, and video technology, including Tim S. Kang, Conor McGill, Camon Crocker, Angel Banchs, Damein Futch, Michael Kohler, and Kris Murray.</p><p class="paragraph" style="text-align:left;">The event follows another PIXERA education initiative that VP Land covered, <a class="link" href="https://www.vp-land.com/p/keynotes-case-studies-and-hands-on-labs-pixera-base-camp-brings-media-server-training-to-nashville?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=hdr-digital-color-in-realtime-video-workflows-pixera-s-santa-monica-educational-summit-focuses-on-seeing-hdr-clearly" target="_blank" rel="noopener noreferrer nofollow">PIXERA Base Camp</a>, a media-server training program. This Santa Monica session is more specific: it is an opportunity to compare perspectives on HDR and digital color in a shared real-time setup.</p><h2 class="heading" style="text-align:left;" id="who-can-attend">Who can attend</h2><p class="paragraph" style="text-align:left;">The summit takes place at PIXERA’s Santa Monica studio, 1745 Berkeley Street, Studio 3, from 9 a.m. to 5 p.m. PT. In-room attendance is limited to 60 people, while the full day will also be available on YouTube. A happy hour at Santa Monica Brew Works follows the program.</p><p class="paragraph" style="text-align:left;">For teams working with LED displays, cameras, and color-managed pipelines, the day is less about a single prescribed HDR workflow than about building a clearer basis for decisions. The useful outcome is a better understanding of where HDR behavior is being defined, how it can be observed, and which creative and technical choices need to be discussed together.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Alibaba Previews Qwen3.8-Max, a 2.4-Trillion-Parameter Multimodal Model With Open Weights to Follow</title>
  <description></description>
      <enclosure length="394806" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/557c943e-aad4-4a62-b9ae-7d457b158ad6/unnamed__19_.jpg"/>
  <link>https://newsletter.vp-land.com/p/alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow</guid>
  <pubDate>Sun, 19 Jul 2026 21:08:05 +0000</pubDate>
  <atom:published>2026-07-19T21:08:05Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Alibaba&#39;s Qwen team has opened preview access to <b>Qwen3.8-Max</b>, a <b>2.4-trillion-parameter</b> model the company describes as its most capable system yet, with open weights promised at an unspecified later date. The announcement came through Qwen&#39;s <a class="link" href="https://x.com/alibaba_qwen/status/2078759124914098291?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow" target="_blank" rel="noopener noreferrer nofollow">official account on X</a>, positioning the release as a multimodal flagship rather than a text-only update.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Available now in preview</b> through Alibaba&#39;s Token Plan, Qoder, and QoderWork.</p></li><li><p class="paragraph" style="text-align:left;"><b>Multimodal input</b> for images, video, and documents.</p></li><li><p class="paragraph" style="text-align:left;"><b>Open weights &quot;soon,&quot;</b> with no date and no Hugging Face model card posted yet.</p></li></ul><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/alibaba_qwen/status/2078759124914098291?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;"><b>A trillion-parameter model built to read images, video, and documents.</b> Qwen positions Qwen3.8-Max as the team&#39;s first multimodal model above 1 trillion parameters, able to take images, video, and documents as input, <a class="link" href="https://the-decoder.com/alibabas-qwen-takes-on-kimi-k3-with-open-weight-qwen-3-8-says-model-is-second-only-to-fable-5/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow" target="_blank" rel="noopener noreferrer nofollow">according to The Decoder</a>. That input range is the practical detail for production teams: a model that ingests footage and reference documents directly fits review, logging, and analysis workflows a text-only model cannot serve. Alibaba also expects it to outperform the prior Qwen3.7-Max on coding and productivity tasks including full-stack development, data analysis, and office work.</p><p class="paragraph" style="text-align:left;">The video-as-input framing tracks with the pitch behind Moonshot&#39;s Kimi K3, <a class="link" href="https://www.vp-land.com/p/moonshot-s-kimi-k3-makes-video-a-native-input-in-a-2-8-trillion-parameter-open-weights-model?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow" target="_blank" rel="noopener noreferrer nofollow">which we covered</a> as a 2.8-trillion-parameter model released days before. The Decoder reports Alibaba&#39;s release is likely aimed at Kimi K3&#39;s momentum.</p><p class="paragraph" style="text-align:left;"><b>&quot;Second only to Fable 5&quot; is a vendor claim, not a benchmark.</b> Qwen says the model matches leading frontier systems and trails only Fable 5, but the company has published no benchmark results to support the ranking. Alibaba has also not disclosed the active-parameter count or the mixture-of-experts configuration, so the 2.4-trillion figure describes total size, not the compute used per query. Without published numbers or third-party testing, the positioning stands as a company claim.</p><p class="paragraph" style="text-align:left;"><b>Preview access runs at a tenth of standard pricing.</b> The Decoder reports the preview is available through Token Plan, Qoder, and QoderWork at 10 percent of the standard price, giving developers a low-cost window to test the model before general availability. Qwen called the preview &quot;continuously evolving,&quot; which means the version teams test today may not match the eventual released weights.</p><p class="paragraph" style="text-align:left;"><b>Open weights would extend Alibaba&#39;s release pattern, but they aren&#39;t here yet.</b> Alibaba has repeatedly shipped models to a paid preview first and released open weights later. The same sequence played out with <a class="link" href="https://www.vp-land.com/p/alibaba-s-wan2-5-preview-breaks-sync-audio-barrier-in-video-generation?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow" target="_blank" rel="noopener noreferrer nofollow">Wan 2.5</a>, which launched as a paid API with open weights promised down the line.</p><p class="paragraph" style="text-align:left;">The company&#39;s open-source track record is real, from the freely downloadable <a class="link" href="https://www.vp-land.com/p/qwen-edit-the-free-ai-image-editor-rivaling-flux-kontext-plus-this-week-s-runway-audio-and-filmmakin?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow" target="_blank" rel="noopener noreferrer nofollow">Qwen Image Edit</a> to a run of open model families across image and video. Qwen3.8-Max stays preview-only for now, and the open-weights claim rests on the announcement alone.</p><p class="paragraph" style="text-align:left;"><b>The open-weights promise is the part still missing.</b> The usable facts are narrow: a multimodal model in paid preview, priced low, accepting image, video, and document input, with no independent benchmarks. The open weights that matter most for self-hosting and fine-tuning, the same draw behind <a class="link" href="https://www.vp-land.com/p/thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=alibaba-previews-qwen3-8-max-a-2-4-trillion-parameter-multimodal-model-with-open-weights-to-follow" target="_blank" rel="noopener noreferrer nofollow">Thinking Machines&#39; Inkling</a> and other open releases, remain a promise without a date. Developers can test the preview today; teams planning around open weights should wait for the model card and real numbers.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Disney Opens the El Capitan and Its Studio Lot to Creators for Created in LA</title>
  <description></description>
      <enclosure length="415544" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d2e603f4-a5b8-447b-a131-60b0275e2735/unnamed__18_.jpg"/>
  <link>https://newsletter.vp-land.com/p/disney-opens-the-el-capitan-and-its-studio-lot-to-creators-for-created-in-la</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/disney-opens-the-el-capitan-and-its-studio-lot-to-creators-for-created-in-la</guid>
  <pubDate>Sun, 19 Jul 2026 21:02:26 +0000</pubDate>
  <atom:published>2026-07-19T21:02:26Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Industry News]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Disney is putting digital creators on a Hollywood red carpet and opening its Burbank studio lot for a new two-day gathering called Created in LA. The event, which Disney describes as its first dedicated creator-focused event, is set for September 17 and 18 and will bring together roughly 350 creators, storytellers, and executives.</p><p class="paragraph" style="text-align:left;">The news has more weight than the partnership announcement that first circulated on LinkedIn. In its <a class="link" href="https://thewaltdisneycompany.com/news/jon-youshaei-created-in-la-event/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=disney-opens-the-el-capitan-and-its-studio-lot-to-creators-for-created-in-la" target="_blank" rel="noopener noreferrer nofollow">official announcement</a>, Disney lays out a format built around two pieces of company real estate that carry real symbolic value: the El Capitan Theatre in Hollywood and the Walt Disney Studios lot in Burbank.</p><p class="paragraph" style="text-align:left;">Created in LA is conceived and hosted with creator and journalist Jon Youshaei. The event is not a film market or a conventional studio premiere. It is Disney creating a staged point of contact between its entertainment infrastructure and an audience that increasingly produces, distributes, and finances its own work.</p><h2 class="heading" style="text-align:left;" id="a-creator-premiere-at-the-el-capita">A creator premiere at the El Capitan</h2><p class="paragraph" style="text-align:left;">The first day takes place at the El Capitan Theatre, where Disney says creator videos will receive the venue&#39;s first red-carpet premiere. That is the clearest signal in the announcement: a theater long associated with Disney’s own tentpole releases will be used to present work made for creator audiences.</p><p class="paragraph" style="text-align:left;">The second day moves to the Walt Disney Studios lot. Disney says the Burbank program will include talks, performances, and keynotes, bringing creators into the same physical setting that has long been reserved for the company’s film and television operation.</p><p class="paragraph" style="text-align:left;">The basics:</p><ul><li><p class="paragraph" style="text-align:left;">September 17-18, 2026</p></li><li><p class="paragraph" style="text-align:left;">Roughly 350 creators, storytellers, and executives</p></li><li><p class="paragraph" style="text-align:left;">A creator-video red-carpet premiere at the El Capitan on day one</p></li><li><p class="paragraph" style="text-align:left;">Talks, performances, and keynotes on the Disney lot on day two</p></li></ul><p class="paragraph" style="text-align:left;">Attendance is application-based through <a class="link" href="https://created.la/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=disney-opens-the-el-capitan-and-its-studio-lot-to-creators-for-created-in-la" target="_blank" rel="noopener noreferrer nofollow">Created in LA</a>. Disney says applications are reviewed on a first-come, first-served basis, rather than positioning the gathering as an open public festival.</p><h2 class="heading" style="text-align:left;" id="disney-is-making-a-physical-bet-on-">Disney is making a physical bet on the creator economy</h2><p class="paragraph" style="text-align:left;">Hollywood studios have spent years trying to figure out where independent creators fit: as marketing partners, as talent pipelines, as competitors for attention, or as future production partners. Created in LA does not settle that question. It does show Disney treating creators as an audience worth convening inside its own venues, with executives and established storytellers in the room.</p><p class="paragraph" style="text-align:left;">That distinction matters. Creator-economy announcements often arrive as software launches, ad-sales initiatives, or content deals. Disney’s choice here is experiential and highly visible. The company is lending the El Capitan&#39;s premiere ritual and the studio lot&#39;s institutional cachet to work that usually premieres on social platforms or creator-owned channels.</p><p class="paragraph" style="text-align:left;">It also comes as Disney is exploring new ways to engage audiences around its intellectual property and digital creation tools. VP Land previously covered <a class="link" href="https://www.vp-land.com/p/disney-planning-genai-content-creation-tools?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=disney-opens-the-el-capitan-and-its-studio-lot-to-creators-for-created-in-la" target="_blank" rel="noopener noreferrer nofollow">Disney’s plan to explore generative-AI creation tools for Disney+ subscribers</a>. Created in LA is a separate initiative, but it fits the same broader question: how much of the next entertainment ecosystem gets built with, rather than simply distributed to, digital creators.</p><h2 class="heading" style="text-align:left;" id="more-than-a-onenight-brand-activati">More than a one-night brand activation</h2><p class="paragraph" style="text-align:left;">The event has enough structure to be more than a red-carpet photo opportunity. A premiere, a studio-lot program, and a capped group of 350 people give Disney a setting for relationship-building across creators, storytellers, and executives. That could lead to little beyond a successful inaugural event, but it could also become a repeatable model for how a major studio identifies talent and tests new creator-facing partnerships.</p><p class="paragraph" style="text-align:left;">For now, Disney has announced a date, venues, format, and application process. The more important story is the invitation itself: the company is bringing creator work into one of Hollywood’s most traditional premiere spaces, then bringing the people behind it onto the Disney lot the next day.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Google Open-Sources GNM Head, a Parametric 3D Head Model With 250+ Identity Controls</title>
  <description></description>
      <enclosure length="322000" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/cebecb16-6dbd-4cb0-a8d5-d22d17f9449e/unnamed__17_.jpg"/>
  <link>https://newsletter.vp-land.com/p/google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls</guid>
  <pubDate>Sun, 19 Jul 2026 20:58:24 +0000</pubDate>
  <atom:published>2026-07-19T20:58:24Z</atom:published>
    <category><![CDATA[Software]]></category>
    <category><![CDATA[Article]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Google has open-sourced <a class="link" href="https://github.com/google/GNM?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls" target="_blank" rel="noopener noreferrer nofollow">GNM Head</a>, a parametric 3D statistical model of the human head that gives artists slider-level control over facial identity and expression. The release is the first piece of GNM (Generative aNthropomorphic Model, pronounced &quot;genome&quot;), which Google describes as an open ecosystem of parametric human models.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Apache 2.0 license.</b> Studios can ship GNM Head inside commercial pipelines and products, not only research projects.</p></li><li><p class="paragraph" style="text-align:left;"><b>636 total parameters.</b> 253 control identity and 383 control expression, down to individual eyes, teeth, and the tongue.</p></li><li><p class="paragraph" style="text-align:left;"><b>Built from scans, not generated.</b> The model is learned from a large dataset of real 3D scans, so the heads are grounded in measured anatomy rather than AI generation.</p></li></ul><p class="paragraph" style="text-align:left;">Google Staff Technical Artist Iker J. de los Mozos, a former Lead Character TD at Walt Disney Animation Studios, announced the release. &quot;My team at Google has just open sourced GNM, our parametric 3D statistical model of the human head, which is learned from a large dataset of 3D scans,&quot; de los Mozos wrote.</p><h2 class="heading" style="text-align:left;" id="253-identity-controls-and-383-expre">253 identity controls and 383 expression controls reach inside the face</h2><p class="paragraph" style="text-align:left;">GNM Head splits the head into two parameter sets. Identity covers the base mesh with 253 controls: 170 for the head, 80 for the teeth, and 3 for the eyeballs. Expression adds 383 more, broken into 100 each for the left and right eyes, 150 for the lower face, 32 for the tongue, and 1 for the irises.</p><p class="paragraph" style="text-align:left;">The model generates male and female heads across a range of ethnicities, and it exposes internal anatomy that most face rigs leave out. Controllable eyeballs, teeth, and a tongue matter for close-up character work, dialogue, and any shot where the inside of the mouth reads on camera.</p><p class="paragraph" style="text-align:left;">This is a statistical model, not a generative AI tool. It produces heads by sampling learned parameters from scan data, which keeps the output anatomically consistent and directly editable instead of hallucinated.</p><h2 class="heading" style="text-align:left;" id="apache-20-and-fourframework-support">Apache 2.0 and four-framework support make it pipeline-ready</h2><p class="paragraph" style="text-align:left;">The license is the part that changes who can use it. GNM Head ships under Apache 2.0, which covers both non-commercial and commercial use, so a studio or a tool vendor can build it into a shipping product without a research-only restriction.</p><p class="paragraph" style="text-align:left;">It also supports the frameworks technical artists and researchers already work in: NumPy, JAX, PyTorch, and TensorFlow, plus semantic parameter sampling. That range lets teams drop the model into existing character or perception pipelines rather than rebuilding around it.</p><p class="paragraph" style="text-align:left;">Meta made a comparable move when it open-sourced <a class="link" href="https://www.vp-land.com/p/meta-s-sam-3d-converts-photos-to-editable-3d-models?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls" target="_blank" rel="noopener noreferrer nofollow">MHR, its parametric human model</a>, alongside SAM 3D. GNM Head narrows that idea to the head and pushes far deeper on facial control.</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="a-free-blender-importer-already-put">A free Blender importer already puts GNM Head in artists&#39; hands</h2><p class="paragraph" style="text-align:left;">Community tooling arrived with the model. Nathan Dickson, a rigging lead whose credits include the Student Academy Award-winning short &quot;Student Accomplice,&quot; released a free GNM Head Importer for Blender 5.1 and up. The add-on exposes sliders for ethnicity, gender, and expression with a live viewport update.</p><p class="paragraph" style="text-align:left;">A Houdini path is in progress as well. David Eschrich, Global Head of 3D at Zoic Studios, built a GNM Head digital asset, though it was not publicly available at launch.</p><p class="paragraph" style="text-align:left;">Open-source graphics tooling like this has become a recurring thread; we covered the <a class="link" href="https://www.vp-land.com/p/dreamworks-moonray-keynote-headlines-aswf-open-source-days-in-la?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls" target="_blank" rel="noopener noreferrer nofollow">ASWF Open Source Days schedule</a>, where Blender and OpenPBR work shared the stage.</p><h2 class="heading" style="text-align:left;" id="the-head-is-the-first-model-with-a-">The head is the first model, with a full human suite on the roadmap</h2><p class="paragraph" style="text-align:left;">GNM Head is the starting point, not the whole plan. Google positions GNM as an open ecosystem of parametric human models and perception stacks, and the project describes a roadmap that &quot;includes releasing a comprehensive suite of statistical models complemented by perception and analysis technology.&quot; Google Senior Research Scientist Stylianos Ploumpis confirmed further releases are coming but said the company could not yet say what or when.</p><p class="paragraph" style="text-align:left;">Google was direct about the model&#39;s limits. The current training data uses binary gender categories and four broad demographic groups, and the company acknowledged it &quot;does not fully represent the spectrum of gender identities or the full diversity of the global population,&quot; with a detailed report promised.</p><p class="paragraph" style="text-align:left;">For VFX and virtual production teams, an Apache-licensed, scan-based head model with this depth of control is a usable building block for digital humans, background crowds, and avatar systems, and it is available on <a class="link" href="https://github.com/google/GNM?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=google-open-sources-gnm-head-a-parametric-3d-head-model-with-250-identity-controls" target="_blank" rel="noopener noreferrer nofollow">Google&#39;s GitHub repository</a>.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Moonshot’s Kimi K3 Makes Video a Native Input in a 2.8-Trillion-Parameter Open-Weights Model</title>
  <description></description>
      <enclosure length="855858" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/63e3b7fb-2d27-45c1-9043-d32fabdee7ff/recjsGCg8fukxRkgU.jpg"/>
  <link>https://newsletter.vp-land.com/p/moonshot-s-kimi-k3-makes-video-a-native-input-in-a-2-8-trillion-parameter-open-weights-model</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/moonshot-s-kimi-k3-makes-video-a-native-input-in-a-2-8-trillion-parameter-open-weights-model</guid>
  <pubDate>Sat, 18 Jul 2026 05:50:29 +0000</pubDate>
  <atom:published>2026-07-18T05:50:29Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Moonshot AI’s <a class="link" href="https://x.com/Kimi_Moonshot/status/2077830229968683203?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=moonshot-s-kimi-k3-makes-video-a-native-input-in-a-2-8-trillion-parameter-open-weights-model" target="_blank" rel="noopener noreferrer nofollow">Kimi K3</a> is a 2.8-trillion-parameter mixture-of-experts model with a 1-million-token context window and native multimodal input. The important detail for production teams is not only that it can inspect images. Kimi K3 can accept video alongside text and images, letting a prompt address the footage itself instead of a transcript or a manually selected set of frames.</p><p class="paragraph" style="text-align:left;"><b>Kimi K3 treats video as vision content inside a multimodal prompt.</b> A team uploads a clip, references it in the request, and pairs it with a text instruction such as identifying a visual event, checking continuity, or locating a moment that matches a brief. <b>Moonshot says it will release K3’s weights on July 27, 2026,</b> opening a self-hosting and fine-tuning path for teams that need more control over their material.</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/bn0atstgavo" width="100%"></iframe><h2 class="heading" style="text-align:left;" id="video-is-an-input-to-the-model-not-">Video is an input to the model, not just a source for a transcript</h2><p class="paragraph" style="text-align:left;">Moonshot’s <a class="link" href="https://platform.kimi.ai/docs/guide/use-kimi-vision-model?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=moonshot-s-kimi-k3-makes-video-a-native-input-in-a-2-8-trillion-parameter-open-weights-model" target="_blank" rel="noopener noreferrer nofollow">vision-model documentation</a> describes multimodal requests as a sequence of text and visual content. For video, a team uploads the clip and references it in the request through a file ID or <code>video_url</code>, alongside the instruction. The model samples key frames from that video and can return text about the visual material it sees.</p><p class="paragraph" style="text-align:left;">That distinction matters in a media workflow. A transcript can tell a model what was said. A contact sheet can show a handful of selected moments. A video request gives K3 sampled key frames from the clip, so it can assess visual material across a sequence rather than only a manually selected still-image set. It is not a frame-by-frame reading of every moment, but it can connect a visual event, an action, or a screen graphic to the prompt without first reducing the entire clip to text.</p><p class="paragraph" style="text-align:left;">The company also pairs that capability with a 1-million-token context window. In practice, the useful question is not whether a facility can fit an entire production in one prompt. It is whether a clip, its related notes, a transcript, and a detailed review instruction can be kept together without reducing the visual material to text before analysis.</p><h2 class="heading" style="text-align:left;" id="where-the-comparison-with-claude-op">Where the comparison with Claude Opus is useful</h2><p class="paragraph" style="text-align:left;">Claude Opus can analyze image inputs, and its <a class="link" href="https://platform.claude.com/docs/en/build-with-claude/vision?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=moonshot-s-kimi-k3-makes-video-a-native-input-in-a-2-8-trillion-parameter-open-weights-model" target="_blank" rel="noopener noreferrer nofollow">vision documentation</a> describes image content blocks for that work. It does not document a native video-input block in the same way. A Claude workflow that needs to reason about footage therefore starts by converting the material into inputs the model accepts, such as extracted frames and a transcript.</p><p class="paragraph" style="text-align:left;">Kimi K3 does not make those workflow artifacts obsolete. A transcript remains useful for dialogue search, and selected frames remain useful for precise visual review. Its difference is that the video file can be part of the model request, with the model sampling key frames rather than relying only on a preselected still-image set. That makes K3 a model to test for logging footage, reviewing a clip against a brief, or combining visual questions with production notes.</p><p class="paragraph" style="text-align:left;">This is a capability comparison, not a blanket quality claim. The right model still depends on the task, the material, the evaluation criteria, and whether a team can run the model reliably on its own infrastructure. Moonshot&#39;s public materials identify video as an input modality, but production teams should test accuracy on their own footage before relying on it for editorial, compliance, or asset-management decisions.</p><h2 class="heading" style="text-align:left;" id="a-28-trillionparameter-model-built-">A 2.8-trillion-parameter model built around long context</h2><p class="paragraph" style="text-align:left;">Kimi K3 uses a mixture-of-experts architecture, which activates only part of its total parameter count for a given request. Moonshot says the model has 2.8 trillion parameters, 1 million tokens of context, always-on reasoning, and native multimodality. It also highlights Kimi Delta Attention, which it says improves decoding speed at long context lengths, and Attention Residuals, which it says improve training efficiency.</p><p class="paragraph" style="text-align:left;">Those are Moonshot&#39;s performance claims, and they need independent testing. The immediate operational fact is that the model is available through Kimi&#39;s products and API, while the company says the full open weights will follow on July 27. If that release arrives as described, facilities will be able to assess a very large model with text, image, and video input without being limited to a hosted API.</p><h2 class="heading" style="text-align:left;" id="open-weights-create-a-different-eva">Open weights create a different evaluation path for post teams</h2><p class="paragraph" style="text-align:left;">For a studio or post house, self-hosting does not automatically mean simple or inexpensive deployment. A model at this scale will require substantial infrastructure, and any video-analysis workflow will need careful tests around clip preparation, latency, security, and failure cases. Open weights do, however, make it possible to evaluate, adapt, and operate the model within a team&#39;s own environment rather than sending every request to a third-party endpoint.</p><p class="paragraph" style="text-align:left;">Before deploying K3, post teams should build a test set of representative clips, define the visual questions the model must answer, and compare its sampled-frame results with a transcript-and-contact-sheet workflow. They should also measure latency, infrastructure cost, and the rate of missed visual events before using the model for any consequential production decision.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>NVIDIA's ARDY Generates Controllable Human Motion in Real Time, and Releases the Code and Weights</title>
  <description></description>
      <enclosure length="2633449" type="image/png" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2b721634-d21b-4597-8153-ee3b57597bc7/unnamed.png"/>
  <link>https://newsletter.vp-land.com/p/nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights</guid>
  <pubDate>Sat, 18 Jul 2026 05:44:44 +0000</pubDate>
  <atom:published>2026-07-18T05:44:44Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">NVIDIA&#39;s Spatial Intelligence Lab published <a class="link" href="https://research.nvidia.com/labs/sil/projects/ardy/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights" target="_blank" rel="noopener noreferrer nofollow">ARDY</a>, a framework that generates 3D human motion in real time from text prompts a user can change while the character is still moving. The lab released the code and model weights publicly rather than keeping the system behind a demo.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Streaming, not batch.</b> ARDY synthesizes motion frame by frame while accepting new instructions mid-sequence, so a character can shift from a stealthy walk to a victory dance without stopping to re-render.</p></li><li><p class="paragraph" style="text-align:left;"><b>Two control paths.</b> Direction comes from online text prompts and from kinematic constraints such as root trajectories, full-body keyframes, and end-effector positions.</p></li><li><p class="paragraph" style="text-align:left;"><b>Open release.</b> The code is public on <a class="link" href="https://github.com/nv-tlabs/ardy?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights" target="_blank" rel="noopener noreferrer nofollow">GitHub</a>, the model weights are on Hugging Face, and the work publishes in ACM Transactions on Graphics.</p></li></ul><h2 class="heading" style="text-align:left;" id="realtime-generation-steered-by-live">Real-time generation steered by live text and sparse constraints</h2><p class="paragraph" style="text-align:left;">ARDY, short for Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation, targets applications that cannot wait for an offline render: animation, simulation, and humanoid robotics. &quot;Generating realistic 3D human motions in real-time within interactive applications is key for animation, simulation, and humanoid robotics,&quot; the researchers write.</p><p class="paragraph" style="text-align:left;">The framework accepts two kinds of direction at once. Text prompts drive behaviors the team demonstrates, including a limp, a pick-and-put action, a stealthy walk, a victory dance, a lean-and-peek, and a zombie sit. Kinematic constraints handle spatial precision: root trajectories and waypoints, full-body keyframes, end-effector positions and rotations, or arbitrary combinations of those. The system also supports long-horizon goal reaching, where a constraint is specified beyond the current generation window, and real-time locomotion control through mouse waypoints and keyboard commands.</p><h2 class="heading" style="text-align:left;" id="a-hybrid-representation-and-a-twost">A hybrid representation and a two-stage denoiser</h2><p class="paragraph" style="text-align:left;">The method splits the motion into two representations. It combines explicit global root motion features with latent body embeddings, which the team frames as a way to balance trajectory control against generative efficiency. On top of that sits a two-stage autoregressive transformer denoiser that predicts the root motion first, then conditions the body-motion prediction on that root output.</p><p class="paragraph" style="text-align:left;">Two additional pieces support the streaming behavior. A variable-length history context captures longer-term semantics to improve generation quality, and masked kinematic constraints allow spatiotemporally sparse conditioning that extends past the current window. The generative core follows the diffusion approach that has moved into character work elsewhere; we covered <a class="link" href="https://www.vp-land.com/p/motorica-secures-5m-to-accelerate-ai-powered-animation-technology?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights" target="_blank" rel="noopener noreferrer nofollow">Motorica&#39;s raise</a> to bring generative AI to character animation.</p><h2 class="heading" style="text-align:left;" id="from-onscreen-characters-to-a-physi">From on-screen characters to a physical humanoid</h2><p class="paragraph" style="text-align:left;">ARDY does not stop at rendered figures. The team integrated its output with a Unitree G1 humanoid robot through the SONIC physical tracking policy, using the generated motion to drive interactive robot control. That connection to hardware fits NVIDIA&#39;s broader robotics push; we covered the company&#39;s GTC keynote and its <a class="link" href="https://www.vp-land.com/p/nvidia-s-gtc-keynote-breakdown-in-30-minutes-youtube?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights" target="_blank" rel="noopener noreferrer nofollow">open humanoid robot dataset</a> for training foundation models.</p><p class="paragraph" style="text-align:left;">For production teams, the interactive framing separates ARDY from capture-first pipelines. Where markerless systems reconstruct motion from a performance, ARDY generates it from instructions; we covered <a class="link" href="https://www.vp-land.com/p/motion-without-markers-move-ai-s-gen-2-technology-redefines-capture-for-filmmakers?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights" target="_blank" rel="noopener noreferrer nofollow">Move AI&#39;s Gen 2</a> markerless capture as one point on that spectrum.</p><h2 class="heading" style="text-align:left;" id="published-at-siggraph-with-code-and">Published at SIGGRAPH with code and weights available</h2><p class="paragraph" style="text-align:left;">The paper appears in ACM Transactions on Graphics, Volume 45, Issue 4, Article 86, presented at SIGGRAPH, with DOI 10.1145/3811284. The authors are Kaifeng Zhao (NVIDIA, ETH Zürich), Mathis Petrovich (NVIDIA), Haotian Zhang (NVIDIA), Tingwu Wang (NVIDIA), Siyu Tang (ETH Zürich), and Davis Rempe (NVIDIA). The <a class="link" href="https://huggingface.co/collections/nvidia/ardy?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=nvidia-s-ardy-generates-controllable-human-motion-in-real-time-and-releases-the-code-and-weights" target="_blank" rel="noopener noreferrer nofollow">model weights</a> are posted on Hugging Face alongside the code.</p><p class="paragraph" style="text-align:left;">Because the code and weights are public, studios and robotics teams can test real-time, promptable motion generation against their own constraints instead of waiting for a product. The open release lets teams check how the models hold up on custom rigs and live control loops directly, rather than through the paper&#39;s demo clips.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Netflix's 300 AI titles</title>
  <description>PLUS: Thinking Machines opens Inkling, playable AI worlds you can steer</description>
      <enclosure length="508688" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0cf946d2-d471-49b3-9728-afae72cf3e6e/telegram-cloud-photo-size-1-5102905422850493981-w.jpg"/>
  <link>https://newsletter.vp-land.com/p/netflix-s-300-ai-titles-67f8</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/netflix-s-300-ai-titles-67f8</guid>
  <pubDate>Sat, 18 Jul 2026 01:33:34 +0000</pubDate>
  <atom:published>2026-07-18T01:33:34Z</atom:published>
    <category><![CDATA[Newsletter]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><hr class="content_break"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0851b96e-9574-41fd-a066-edbf06c8844c/VP_Land_Newsletter_Banner_-_18.png"/></div><hr class="content_break"><div class="section" style="background-color:#FFFFFF;border-color:#c75a5e;border-radius:50px;border-style:solid;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;"><b>Welcome to VP Land!</b> It&#39;s SIGGRAPH this week! Tons of official (and unofficial) events - if you&#39;ll be around just reply back.</p><p class="paragraph" style="text-align:left;"><b>One thing to put on your radar - a free summit Tuesday night on some groundbreaking </b><b><a class="link" href="https://www.eventbrite.com/e/ai-workflows-summit-tickets-1993363191967?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">AI Workflows</a></b><b> (aptly called the AI Workflows Summit).</b> It&#39;s a night focused on how AI is actually getting used in professional production, from virtual production and volumetric media to the infrastructure and tools behind the workflows, including ComfyUI. A few in-depth presentations, followed by a panel that I&#39;ll be moderating. Tuesday from 6PM-9:30PM, completely free, even if you don&#39;t have a SIGGRAPH badge. Register <a class="link" href="https://www.eventbrite.com/e/ai-workflows-summit-tickets-1993363191967?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">here</a>.</p><p class="paragraph" style="text-align:left;">In today&#39;s edition:</p><ul><li><p class="paragraph" style="text-align:left;">Netflix&#39;s roughly 300 AI-assisted titles</p></li><li><p class="paragraph" style="text-align:left;">Moonshot&#39;s 2.8T open-weights Kimi K3</p></li><li><p class="paragraph" style="text-align:left;">Thinking Machines opens up Inkling</p></li><li><p class="paragraph" style="text-align:left;">Foundry&#39;s SmartRoto and Griptape ship</p></li><li><p class="paragraph" style="text-align:left;">Open world models you can actually steer</p></li></ul><p class="paragraph" style="text-align:left;"><b>One more thing:</b> I have two tickets to The Odyssey in 70mm IMAX, 11:10AM Sunday at the Regal Spectrum in Irvine, that I can&#39;t use (SIGGRAPH, whoops). If you can use them yourself (personal use only, I&#39;ll hunt you down if you resell them), reply to this email. If more than one person can use them, whoever has sent the most <a class="link" href="{{rp_refer_url}}" target="_blank" rel="noopener noreferrer nofollow">newsletter referrals</a> is the tie breaker.</p></div><hr class="content_break"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1a8ef6ca-1f50-47f7-bb27-bde9ab288e6c/Video_Tape_-_Box_Small.png?t=1712408310"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-width:0px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">Netflix Used Generative AI in Roughly 300 Titles in 2026, Concentrated in Post</h3></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1287c049-63a2-4af2-896b-80713870e2d1/unnamed__16_.jpg?t=1784325491"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">Netflix told shareholders in its <a class="link" href="https://variety.com/2026/biz/news/about-300-netflix-programs-used-ai-this-year-q2-earnings-1236812914/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Q2 earnings report</a> that roughly 300 of its titles used generative AI in 2026, the clearest number it has attached to production-side AI, with most of it landing in post-production.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Where the AI actually shows up.</b> Netflix said the tools span concept, pre-visualization, post-production, and release, but the volume sits in post, from cleanup to complex shot work rather than fully generated scenes.</p></li><li><p class="paragraph" style="text-align:left;"><b>What it built on screen.</b> The company pointed to the Indian sports thriller &quot;Glory,&quot; the Brazilian soccer miniseries &quot;Brasil 70: A Saga do Tri,&quot; and the Revolution-era docuseries &quot;The American Experiment,&quot; where AI helped assemble highly complex sequences with enhanced crowd sizes and battle scenes.</p></li><li><p class="paragraph" style="text-align:left;"><b>Sarandos leads with quality, not cost.</b> Co-CEO Ted Sarandos framed the output as &quot;10% better&quot; creatively, and Netflix said the tools let creators expand a project&#39;s scope &quot;twice as fast and at half the cost of previous options.&quot;</p></li><li><p class="paragraph" style="text-align:left;"><b>The figure sits on a growing in-house stack.</b> We covered Netflix&#39;s <a class="link" href="https://www.vp-land.com/p/netflix-releases-its-first-public-ai-model-and-it-s-built-for-post?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">first public model</a>, VOID, built for post-production fixes; the 300-title number shows how far that post-first approach has spread across the slate.</p></li></ul></div><hr class="content_break"><div class="section" style="background-color:transparent;border-color:#3a8acc;border-radius:50px;border-style:solid;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:center;"><span style="color:#0194f6;"><b>SPONSOR MESSAGE</b></span></p><h3 class="heading" style="text-align:left;">Your prompts are leaving out 80% of what you&#39;re thinking.</h3><div class="image"><a class="image__link" href="https://ref.wisprflow.ai/beehiiv-ai/?utm_campaign={{publication_alphanumeric_id}}&utm_source=beehiiv&utm_term=ai_p1_q3&_bhiiv=opp_cd46feb6-fd34-427d-9056-808c58c1de45_4de8c0ec&bhcl_id=93ca25fd-1b18-459f-be2e-93f7282c1d89_{{subscriber_id}}_{{email_address_id}}" rel="noopener" target="_blank"><img class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a2f9e24a-0a18-4e00-8571-aced30127187/flow-89-percent-no-edits.png?t=1776897816"/></a></div><p class="paragraph" style="text-align:left;">When you type a prompt, you summarize. When you speak one, you explain. <a class="link" href="https://ref.wisprflow.ai/beehiiv-ai/?utm_campaign={{publication_alphanumeric_id}}&utm_source=beehiiv&utm_term=ai_p1_q3&_bhiiv=opp_cd46feb6-fd34-427d-9056-808c58c1de45_4de8c0ec&bhcl_id=93ca25fd-1b18-459f-be2e-93f7282c1d89_{{subscriber_id}}_{{email_address_id}}" target="_blank" rel="noopener noreferrer nofollow">Wispr Flow</a> captures your full reasoning — constraints, edge cases, examples, tone — and turns it into clean, structured text you paste into ChatGPT, Claude, or any AI tool. The difference shows up immediately. More context in, fewer follow-ups out.</p><p class="paragraph" style="text-align:left;">89% of messages sent with zero edits. Used by teams at OpenAI, Vercel, and Clay. Try Wispr Flow free — works on Mac, Windows, and iPhone.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://ref.wisprflow.ai/beehiiv-ai/?utm_campaign={{publication_alphanumeric_id}}&utm_source=beehiiv&utm_term=ai_p1_q3&_bhiiv=opp_cd46feb6-fd34-427d-9056-808c58c1de45_4de8c0ec&bhcl_id=93ca25fd-1b18-459f-be2e-93f7282c1d89_{{subscriber_id}}_{{email_address_id}}" target="_blank" rel="noopener noreferrer nofollow">Start flowing free</a></p></div><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><hr class="content_break"></div><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/845ebb56-17b5-4d6f-9526-5e36c53ed637/Light_-_Box_Small.png?t=1720210827"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-width:0px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">Moonshot&#39;s Kimi K3 Lands as the Largest Open-Weights Model Yet, at 2.8 Trillion Parameters</h2></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/6ee65a6d-1e5f-422c-9ef5-d16034b7e2b9/second-unit-kimi.jpg?t=1784325642"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">Moonshot AI released <a class="link" href="https://x.com/Kimi_Moonshot/status/2077830229968683203?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Kimi K3</a>, a 2.8-trillion-parameter mixture-of-experts model it calls the largest open-weights model available, with a 1-million-token context window and native visual understanding.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Frontier-class, and downloadable.</b> Kimi K3 runs an always-on reasoning mode and posts near-frontier benchmark scores; Moonshot says it will release the full weights on July 27, keeping a model this size self-hostable for teams that want to fine-tune it.</p></li><li><p class="paragraph" style="text-align:left;"><b>Built to scale around compute limits.</b> The model uses Moonshot&#39;s Kimi Delta Attention and Attention Residuals, and is roughly 75% larger than DeepSeek&#39;s V4 Pro, extending China&#39;s run of heavyweight open releases.</p></li><li><p class="paragraph" style="text-align:left;"><b>Video in, analysis out.</b> K3 accepts video alongside text and images, so a clip can be described, logged, or checked against a brief without first being reduced to a transcript.</p></li></ul></div><hr class="content_break"><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:0px;border-bottom-right-radius:0px;border-bottom-width:0px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:50px;border-top-right-radius:50px;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">Thinking Machines Releases Inkling, Its First Open-Weights Model Built to Be Fine-Tuned</h2></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/96784002-10bc-46f0-b70e-50b0ffd75e55/extra-box-inkling.jpg?t=1784325673"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#ffb914;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">Thinking Machines Lab, the San Francisco company founded by former OpenAI CTO Mira Murati, released <a class="link" href="https://thinkingmachines.ai/news/introducing-inkling/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Inkling</a>, its first in-house model and an open-weight foundation for teams that need more control over how a model is run.</p><ul><li><p class="paragraph" style="text-align:left;"><b>A model facilities can adapt.</b> Inkling&#39;s weights are available under Apache 2.0, so a studio can fine-tune it on its own material through the company&#39;s Tinker platform and host the result inside its network.</p></li><li><p class="paragraph" style="text-align:left;"><b>Multimodal training, text output.</b> The 975-billion-parameter mixture-of-experts model was trained on text, images, audio, and video, with text-only outputs at launch.</p></li><li><p class="paragraph" style="text-align:left;"><b>A US-based open option.</b> Inkling gives creators and facilities another open model to evaluate alongside releases from Chinese labs, while Thinking Machines positions it as a broad generalist rather than a benchmark leader.</p></li></ul><p class="paragraph" style="text-align:left;">Open weights and in-house fine-tuning make it possible to adapt a capable model to a facility&#39;s footage while keeping that material on its own servers.</p></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e619d014-9e24-4878-b569-04b3b2d62bb8/Toolbox_-_Box_Small.png?t=1738938789"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-width:0px;border-color:#3a8acc;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">From Roto to Real-Time Worlds</h2></div><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d9020122-160a-44c2-bcd4-fc8927e83128/ai-tools.jpg?t=1784325687"/></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-bottom-width:4px;border-color:#3a8acc;border-left-width:4px;border-right-width:4px;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-top-width:0px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;">This tool slate is less about watching polished demos and more about putting controllable AI into a working pipeline: local roto, studio-safe orchestration, steerable worlds, and models built for revision.</p><h3 class="heading" style="text-align:left;">Production pipeline</h3><ul><li><p class="paragraph" style="text-align:left;"><b>Griptape Enterprise.</b> Foundry&#39;s <a class="link" href="https://www.foundry.com/news-and-awards/foundry-announces-griptape-enterprise?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">studio platform</a> runs agentic AI pipelines on-premises or in a private cloud, with licensing, permissions, Nuke integration, and modular diffusion workflows for shops that cannot send footage to a public API.</p></li></ul><ul><li><p class="paragraph" style="text-align:left;"><b>SmartRoto for Nuke.</b> Foundry&#39;s AI add-on <a class="link" href="https://www.foundry.com/insights/machine-learning/smartroto-enabling-rotoscoping?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">moves and deforms roto splines across a shot</a> from a few input frames while keeping the artist in control. It runs locally on licensed data, and Foundry says it can cut rotoscoping time by up to four times.</p></li></ul><ul><li><p class="paragraph" style="text-align:left;"><b>Mirelo AI and Kyutai audio-to-MIDI.</b> This open model <a class="link" href="https://www.mirelo.ai/blog/turning-audio-to-midi?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">transcribes a full mix</a> into a separate MIDI track per instrument, where many tools handle one sound at a time. It gives music editorial and scoring teams editable parts instead of a locked stereo mix.</p></li></ul><h3 class="heading" style="text-align:left;">Directable worlds and interactive media</h3><ul><li><p class="paragraph" style="text-align:left;"><b>Lingbot-World-2.</b> Robbyant&#39;s open-source world model, hosted <a class="link" href="https://x.com/reactorworld/status/2074935025771143626?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">on Reactor</a>, runs in real time and lets users steer a character and camera by keyboard. Its persistent environments and available API point toward previz and virtual-location exploration.</p></li></ul><ul><li><p class="paragraph" style="text-align:left;"><b>Aval.</b> Alex Barashkov&#39;s <a class="link" href="https://x.com/alex_barashkov/status/2077406669647073370?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">open interactive-video format</a> brings state machines, frame-accurate transitions, and alpha transparency to web playback. It gives teams authored triggers and branching behavior without a game engine.</p></li></ul><ul><li><p class="paragraph" style="text-align:left;"><b>AlayaWorld.</b> Fine-tuned from LTX-2.3, <a class="link" href="https://x.com/aisearchio/status/2074981411212472425?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">AlayaWorld</a> generates playable 720p worlds at 24fps with real-time 6-DoF camera control and prompt changes mid-scene. Spatial memory helps revisited locations remain consistent past the one-minute mark.</p></li></ul><h3 class="heading" style="text-align:left;">Image and research models</h3><ul><li><p class="paragraph" style="text-align:left;"><b>Reve 2.1.</b> The <a class="link" href="https://x.com/reve/status/2075248950756716747?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">native-4K image-model update</a> improves prompt understanding, world knowledge, and foreign-text rendering. That is useful when a generated frame needs legible non-English type rather than placeholder text.</p></li></ul><ul><li><p class="paragraph" style="text-align:left;"><b>NVIDIA ARDY.</b> This <a class="link" href="https://research.nvidia.com/labs/sil/projects/ardy/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">autoregressive diffusion model</a> produces real-time 3D human movement that can be redirected with live text prompts and constrained with waypoints or full-body keyframes, targeting animation, simulation, and robotics.</p></li></ul></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7b3e63fb-94ff-4fda-9feb-6daa731cfb93/Television_-_Box_Small.png?t=1738625951"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#b4d8db;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">How Disney Built a New Smugglers Run Mission Featuring The Mandalorian & Grogu</h3><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/b7GPkRpIy8s" width="100%"></iframe><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/b7GPkRpIy8s?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">How Disney Built a New Smugglers Run Mission</a> goes behind the Galaxy&#39;s Edge update that folds The Mandalorian and Grogu into the ride. The Disney Parks video walks through the ride-design and media work behind the new mission.</p></div><hr class="content_break"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7d8b5c9f-853b-47b9-a096-82a3f91451a6/Clothespin_-_Box_Small.png?t=1738623074"/></div><div class="section" style="background-color:transparent;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#c40101;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><p class="paragraph" style="text-align:left;"><i>Stories, projects, and links that caught our attention from around the web:</i></p><p class="paragraph" style="text-align:left;">🎬 A <a class="link" href="https://www.yahoo.com/entertainment/movies/articles/hollywood-ai-hype-hit-reality-140000812.html?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">studio-adoption survey</a> finds AI landing in previz, scheduling, and sports repackaging even as 96% of CEOs report no clear ROI.</p><p class="paragraph" style="text-align:left;">🔊 A music-industry coalition proposed voluntary <a class="link" href="https://x.com/Variety/status/2075585333119594834?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">streaming labels</a> tagging songs as AI-Generated or AI-Assisted, modeled on the explicit-lyrics marker.</p><p class="paragraph" style="text-align:left;">👓 OpenAI and Work Louder launched <a class="link" href="https://x.com/OpenAIDevs/status/2077425991790870644?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">kbd-1.0-codex-micro</a>, a limited-run macropad with mappable keys and a joystick for the Codex workflow.</p><p class="paragraph" style="text-align:left;">▶️ LTX <a class="link" href="https://x.com/ltx_io/status/2074904819731661101?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">spun out</a> of Lightricks as an independent open world-models company building foundation models for filmmaking, gaming, and robotics.</p></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2bd899e4-ee40-4667-8548-4ba64b6fef8f/Denoised_-_Box_Small_-_01.png?t=1745256407"/></div></div><div class="section" style="background-color:transparent;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#7600c3;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h3 class="heading" style="text-align:left;">We Tested Seedream 5.0 Pro While Meta Backpedaled on Instagram AI Remixes</h3><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/pmuDx9mo-Tw" width="100%"></iframe><p class="paragraph" style="text-align:left;">On the latest <a class="link" href="https://www.youtube.com/watch?v=pmuDx9mo-Tw&utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Denoised</a>, Addy Ghani and Joey work through a dense run of model news for production teams. They break down Seedream 5.0 Pro&#39;s more precise image editing, Meta&#39;s Muse model and its opt-out Instagram remix feature, and what new MCP integrations for Unreal Engine and ComfyUI mean for wiring AI into a real pipeline. The throughline is control: where these tools hand you tighter, repeatable command over output, and where an opt-out default quietly shifts the burden back onto creators.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://Nvidia.Read?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Read</a> the <a class="link" href="https://www.vp-land.com/p/ces-2026-fuji-s-fake-8mm-camera-lego-smart-bricks-atlas-robot-and-more?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">show notes</a> or watch the <a class="link" href="https://www.youtube.com/watch?v=pmuDx9mo-Tw_&utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">full episode</a>.</p><p class="paragraph" style="text-align:left;"><b>Watch/Listen & Subscribe</b></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://vpgo.link/denoised-spotify?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Spotify</a> | <a class="link" href="https://vpgo.link/denoised-apple?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Apple Podcast</a> | <a class="link" href="https://vpgo.link/denoised-youtube?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">YouTube</a></p></div><hr class="content_break"><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1d8019c1-2e72-46e3-a35e-925e89e79f90/Clipboard_-_Box_Small_1_.png?t=1712401303"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#219c30;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">👔 Open Job Posts</h2><p class="paragraph" style="text-align:left;">🆕 <a class="link" href="https://job-boards.eu.greenhouse.io/creativefabrica/jobs/4856784101?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">AI Video Creator & Edito</a>r - Remote<br>Creative Fabrica</p><p class="paragraph" style="text-align:left;">🆕 <a class="link" href="https://jobs.ashbyhq.com/runway-ml/5e8d6c9b-0d30-4bb2-90a9-df7d937aa44e?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Research Science Manager, Foundation Models</a> - Remote<br>Runway</p><p class="paragraph" style="text-align:left;">🆕 <a class="link" href="https://job-boards.greenhouse.io/blackforestlabs?error=true&utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Forward Deployed Machine Learning Engineer</a> - San Francisco, CA<br>Black Forest Labs</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vfxengine.com/jobs/generative-ai-supervisor/907d1bc6-188d-48d7-a90a-87fe4b63b99c?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Generative AI Supervisor</a> - Milan, Italy<br>VFX Engine</p></div><hr class="content_break"><div id="events" class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/13731556-210f-41c3-af6d-0fe92e7099ae/Calendar_-_Box_Small.png?t=1712401007"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#3a8acc;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><h2 class="heading" style="text-align:left;">📆 Upcoming Events</h2><p class="paragraph" style="text-align:left;"><b>Jul 19 to 23</b><br><a class="link" href="https://s2026.siggraph.org/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">SIGGRAPH 2026</a><br>Los Angeles, CA</p><p class="paragraph" style="text-align:left;">🆕 <b>Jul 20 to 22</b><br><a class="link" href="https://www.digitalhollywood.com?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">AI & Entertainment Experience / Super Creativity at Digital Hollywood</a><br>Virtual</p><p class="paragraph" style="text-align:left;"><b>Aug 4 to 6</b><br><a class="link" href="https://ai4.io?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Ai4 2026</a><br>Las Vegas, NV</p><p class="paragraph" style="text-align:left;">🆕 <b>Aug 1 to Sep 5</b><br><a class="link" href="https://pages.becomecgpro.com/ai-for-filmmakers-course-live?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">AI for Filmmakers Live Course (CG Pro)</a><br>Virtual</p><p class="paragraph" style="text-align:left;"><b>Sept 11 to 14</b><br><a class="link" href="https://show.ibc.org?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">IBC 2026</a><br>Amsterdam, Netherlands</p><p class="paragraph" style="text-align:left;"><i>View the full event calendar and submit your own events </i><i><a class="link" href="https://www.vp-land.com/vp-events?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">here</a></i><i>. </i></p></div><hr class="content_break"><div id="section" class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1206dfb7-86ea-4af9-a638-7b591764cf0c/Martini_-_Box_Small.png?t=1738624339"/></div></div><div class="section" style="background-color:#FFFFFF;border-bottom-left-radius:50px;border-bottom-right-radius:50px;border-color:#b4d8db;border-style:solid;border-top-left-radius:0px;border-top-right-radius:0px;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:30.0px 20.0px 30.0px 20.0px;"><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/089b7679-7f7d-4dac-8072-f1ec45e74225/CleanShot_2026-07-17_at_14.53.20_2x.png?t=1784326544"/></div></div><hr class="content_break"><hr class="content_break"><div class="section" style="background-color:transparent;border-color:#c40101;border-radius:50px;border-style:solid;border-width:4px;margin:0.0px 0.0px 0.0px 0.0px;padding:20.0px 20.0px 20.0px 20.0px;"><p class="paragraph" style="text-align:left;"><b>Thanks for reading VP Land!</b></p><p class="paragraph" style="text-align:left;">Thanks for reading VP Land!</p><p class="paragraph" style="text-align:left;">Have a link to share or a story idea? <a class="link" href="https://newterritory.media/contact/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Send it here</a>.</p><p class="paragraph" style="text-align:left;">Interested in reaching media industry professionals? <a class="link" href="https://newterritory.media/advertise-with-vp-land/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-s-300-ai-titles" target="_blank" rel="noopener noreferrer nofollow">Advertise with us</a>.</p></div></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:subtitle>PLUS: Thinking Machines opens Inkling, playable AI worlds you can steer</itunes:subtitle><itunes:author>Coffee and Celluloid</itunes:author><itunes:summary>PLUS: Thinking Machines opens Inkling, playable AI worlds you can steer</itunes:summary><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Aval Brings State-Driven Interactive Video to the Web as an Open Source Format</title>
  <description></description>
      <enclosure length="217364" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fd8f8ba3-6f9a-41f9-82ee-ac24d0fd1911/unnamed__15_.jpg"/>
  <link>https://newsletter.vp-land.com/p/aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format</guid>
  <pubDate>Fri, 17 Jul 2026 05:59:05 +0000</pubDate>
  <atom:published>2026-07-17T05:59:05Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Developer Alex Barashkov has released <b>Aval</b>, an open source format for interactive video on the web that responds to hover, click, and application state instead of playing back as a fixed clip. He <a class="link" href="https://x.com/alex_barashkov/status/2077406669647073370?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format" target="_blank" rel="noopener noreferrer nofollow">announced the technical preview</a> on X.</p><p class="paragraph" style="text-align:left;">The <a class="link" href="https://github.com/pixel-point/aval?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format" target="_blank" rel="noopener noreferrer nofollow">code is on GitHub</a> under an MIT license. Barashkov is CEO of the design studio Pixel Point.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/alex_barashkov/status/2077406669647073370?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;"><b>Aval targets a narrow but persistent problem:</b> short prerendered animations, such as UI icons, that need to react to user input without a developer hand-timing every transition or seeking through a video file at each interaction.</p><p class="paragraph" style="text-align:left;">The format packs three things into one asset:</p><ul><li><p class="paragraph" style="text-align:left;"><b>A deterministic state graph.</b> Named states and authored triggers define behavior, with routing where the latest trigger wins.</p></li><li><p class="paragraph" style="text-align:left;"><b>Frame-accurate transitions.</b> Routes begin on authored content frames using portals, finishes, cuts, and reversals rather than approximate media seeks.</p></li><li><p class="paragraph" style="text-align:left;"><b>Packed alpha transparency.</b> Transparent prerendered motion composites through WebGL2, so animations sit cleanly over any interface.</p></li></ul><h2 class="heading" style="text-align:left;" id="the-format-pairs-a-compiler-with-a-">The format pairs a compiler with a web renderer built on WebCodecs and WebGL2</h2><p class="paragraph" style="text-align:left;">Aval ships in two parts. A compiler produces the asset, and a web renderer plays it in the browser.</p><p class="paragraph" style="text-align:left;">The runtime decodes with WebCodecs and renders with WebGL2, keeping the decoder timeline moving forward across loop seams instead of seeking backward. Barashkov calls these &quot;seekless loops,&quot; and points to low CPU overhead and small file sizes as the payoff, describing the format as suited to small icons designed and animated in Blender.</p><p class="paragraph" style="text-align:left;">For browsers that lack the required APIs, or for users with reduced-motion settings, Aval falls back to host-owned markup that the developer supplies. That progressive fallback keeps the underlying content available when the interactive layer cannot run.</p><h2 class="heading" style="text-align:left;" id="aval-is-an-open-source-answer-to-ai">Aval is an open source answer to Airbnb&#39;s Lava</h2><p class="paragraph" style="text-align:left;">Barashkov framed the project as a response to work he could see but not use. Airbnb built <a class="link" href="https://medium.com/@waldobear002/airbnbs-new-lava-icon-format-a-technical-deep-dive-b2604626c7e0?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format" target="_blank" rel="noopener noreferrer nofollow">a similar format called Lava</a>, a compact 3D animated icon format with transparent backgrounds designed for its app interface.</p><p class="paragraph" style="text-align:left;">&quot;Then, a little over a year ago, Airbnb created Lava for almost exactly the same purpose. That gave me hope,&quot; Barashkov wrote. &quot;But Lava was never released as open source. So I decided to build my own, with AI.&quot;</p><p class="paragraph" style="text-align:left;">For motion designers who build broadcast and interface animation, that positions Aval alongside the tooling we covered in <a class="link" href="https://www.vp-land.com/p/maxon-one-update-targets-motion-graphics-unreal-pipelines?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=aval-brings-state-driven-interactive-video-to-the-web-as-an-open-source-format" target="_blank" rel="noopener noreferrer nofollow">Maxon One&#39;s Cinema 4D and Redshift updates</a>, but aimed at the web runtime rather than the render.</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="built-solo-with-an-ai-coding-agent">Built solo with an AI coding agent</h2><p class="paragraph" style="text-align:left;">&quot;This is probably the craziest thing I&#39;ve ever built with Codex,&quot; Barashkov wrote, referring to OpenAI&#39;s coding agent. He said the format had been a years-long goal that never penciled out before AI-assisted development.</p><p class="paragraph" style="text-align:left;">&quot;Before AI, building it would have taken months of work. I could never justify that investment for a noncommercial open source project,&quot; he wrote.</p><p class="paragraph" style="text-align:left;">A single designer shipped a compiler, a web runtime, and a new asset format, matching a capability that a company the size of Airbnb kept internal. AI-assisted coding is what made that solo effort feasible, according to Barashkov.</p><h2 class="heading" style="text-align:left;" id="an-early-preview-not-a-production-d">An early preview, not a production dependency</h2><p class="paragraph" style="text-align:left;">Aval is a technical preview, and Barashkov said he will keep polishing it in the coming days. Its runtime depends on WebCodecs and WebGL2, so behavior varies by browser, and the fallback path matters for anyone shipping to a broad audience.</p><p class="paragraph" style="text-align:left;">For teams already exporting icon and micro-animation sequences from Blender, the release offers a way to keep those animations interactive and lightweight on the web without waiting for a proprietary format to open up. Whether Aval attracts contributors past the preview will determine if it becomes a dependable option or a well-documented experiment.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Netflix Used Generative AI in Roughly 300 Titles in 2026, Concentrated in Post</title>
  <description></description>
      <enclosure length="399274" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/eb1c0393-ec40-4ce4-91f4-6725b448f292/unnamed__14_.jpg"/>
  <link>https://newsletter.vp-land.com/p/netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post</guid>
  <pubDate>Fri, 17 Jul 2026 05:40:17 +0000</pubDate>
  <atom:published>2026-07-17T05:40:17Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Netflix disclosed that generative AI workflows were used in roughly 300 of its titles in 2026, the clearest number the company has put on its production-side AI use so far. The figure appeared in Netflix&#39;s <a class="link" href="https://www.sec.gov/Archives/edgar/data/0001065280/000106528026000211/ex991_q226.htm?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post" target="_blank" rel="noopener noreferrer nofollow">second-quarter shareholder letter</a>.</p><p class="paragraph" style="text-align:left;"><b>Netflix said the largest concentration of that work sits in post-production.</b> The letter described GenAI as scaling &quot;across the production lifecycle, from concept and pre-visualization through post and delivery,&quot; and named three productions that used the tools for shots the company says would have been cut otherwise. Variety <a class="link" href="https://variety.com/2026/biz/news/about-300-netflix-programs-used-ai-this-year-q2-earnings-1236812914/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post" target="_blank" rel="noopener noreferrer nofollow">first flagged the disclosure</a> inside the earnings report.</p><h2 class="heading" style="text-align:left;" id="gen-ai-now-touches-every-production">GenAI now touches every production stage, with post carrying the most work</h2><p class="paragraph" style="text-align:left;">Netflix told shareholders that &quot;GenAI workflows have been used in roughly 300 of our titles&quot; in 2026, and described adoption among its creative partners as scaling quickly rather than a set of one-off tests. The heaviest use is downstream, in post-production.</p><p class="paragraph" style="text-align:left;">The pitch to investors was about cost and speed. &quot;We are increasingly leveraging these tools to deliver higher quality output more quickly and at a lower cost than traditional methods,&quot; the company wrote. Netflix has <a class="link" href="https://www.vp-land.com/p/netflix-embraces-ai-efficiency-ai-stories-pass-blind-tests-and-coogler-s-landmark-copyright-deal?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post" target="_blank" rel="noopener noreferrer nofollow">made this efficiency case before</a>, telling investors that AI can lift the finished product, not only trim the budget.</p><h2 class="heading" style="text-align:left;" id="three-named-productions-used-gen-ai">Three named productions used GenAI for crowds, battles, and worldbuilding</h2><p class="paragraph" style="text-align:left;">Netflix singled out <b>Glory</b> (India), <b>Brasil 70: A Saga do Tri</b> (Brazil), and <b>The American Experiment</b> (US) as titles that used GenAI to build &quot;highly complex sequences.&quot; The company listed the specific applications as enhanced crowds, historical battle sequences, and worldbuilding establishing shots.</p><p class="paragraph" style="text-align:left;">Netflix tied those shots to access, not just savings. &quot;In some cases, productions would have had to leave out key shots and sequences in the absence of GenAI technology,&quot; the letter said, positioning the tools as a way for productions to reach for scenes that would otherwise sit outside their budgets.</p><h2 class="heading" style="text-align:left;" id="the-300-figure-sits-on-top-of-netfl">The 300 figure sits on top of Netflix&#39;s build-in-house AI strategy</h2><p class="paragraph" style="text-align:left;">Netflix has spent 2026 building its production AI internally rather than licensing general-purpose tools. Co-CEO Ted Sarandos restated the approach on the earnings call, saying the company is &quot;primarily builders, not buyers, and that remains the case today.&quot;</p><p class="paragraph" style="text-align:left;">That strategy is now visible across the pipeline. We covered <a class="link" href="https://www.vp-land.com/p/netflix-releases-its-first-public-ai-model-and-it-s-built-for-post?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post" target="_blank" rel="noopener noreferrer nofollow">Netflix&#39;s first public AI model</a>, VOID, which is built specifically for post-production object removal, the same stage where the company says most of its GenAI work is concentrated.</p><p class="paragraph" style="text-align:left;">The company has also been buying and building teams to feed that pipeline. We covered its acquisition of <a class="link" href="https://www.vp-land.com/p/netflix-acquires-ben-affleck-s-interpositive-bringing-ai-post-production-tools-in-house?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post" target="_blank" rel="noopener noreferrer nofollow">Ben Affleck&#39;s AI startup InterPositive</a>, which pulled post-production tooling in-house.</p><p class="paragraph" style="text-align:left;">On the content side, Netflix stood up <a class="link" href="https://www.vp-land.com/p/netflix-s-inkubator-brings-genai-native-animation-pipelines-to-shorts-and-specials?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=netflix-used-generative-ai-in-roughly-300-titles-in-2026-concentrated-in-post" target="_blank" rel="noopener noreferrer nofollow">INKubator</a>, an internal studio building GenAI-native animation pipelines for short-form content.</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="ai-also-spread-into-discovery-and-t">AI also spread into discovery and the ads business</h2><p class="paragraph" style="text-align:left;">Beyond production, Netflix said it is using large language models to improve title discovery and rolling out voice search and AI-powered natural language search for members. On the advertising side, the company said it expanded AI tools &quot;across the full advertising lifecycle, from planning and creative production to campaign management, optimization, and reporting.&quot;</p><p class="paragraph" style="text-align:left;">That ads push runs alongside a target Netflix reaffirmed of roughly $3 billion in ad revenue for 2026.</p><h2 class="heading" style="text-align:left;" id="gen-ai-shifts-from-pilot-to-standar">GenAI shifts from pilot to standard line item at the largest streamer</h2><p class="paragraph" style="text-align:left;">Netflix placed the 300-title figure in a shareholder letter, next to Q2 revenue of $12.6 billion, up 13% year over year, and a 33% operating margin. Reporting production AI beside its financials, rather than as an experiment, signals how routine the tools have become at the company.</p><p class="paragraph" style="text-align:left;">For post houses and international productions, GenAI is already inside hundreds of shipped Netflix titles, weighted toward post, and Netflix is now tracking that work as a standard part of how its shows get made.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Thinking Machines Releases Inkling, Its First Open-Weights Model Built to Be Fine-Tuned</title>
  <description></description>
      <enclosure length="365637" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fc1cc8c0-3dd7-4889-9a64-d5ce253e369e/unnamed__13_.jpg"/>
  <link>https://newsletter.vp-land.com/p/thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned</guid>
  <pubDate>Fri, 17 Jul 2026 05:36:10 +0000</pubDate>
  <atom:published>2026-07-17T05:36:10Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Thinking Machines Lab released <a class="link" href="https://thinkingmachines.ai/news/introducing-inkling/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned" target="_blank" rel="noopener noreferrer nofollow">Inkling</a>, its first in-house AI model, and put the full weights up for download. It is an open-weights, multimodal Mixture-of-Experts model that takes text, images, audio, and video as input, and the company is positioning it less as a finished product than as a base other teams retrain for their own work.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Open weights, downloadable now.</b> The full model weights are on Hugging Face in both standard and NVFP4 formats, the latter tuned for NVIDIA Blackwell hardware.</p></li><li><p class="paragraph" style="text-align:left;"><b>Two sizes.</b> Inkling runs 975 billion total parameters with 41 billion active per token. A preview variant, Inkling-Small, runs 276 billion total and 12 billion active.</p></li><li><p class="paragraph" style="text-align:left;"><b>Built to be customized.</b> Thinking Machines is steering users toward fine-tuning Inkling through Tinker, its model-customization platform, rather than shipping it as a one-size-fits-all system.</p></li></ul><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/thinkymachines/status/2077454609551921208?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned"><p> Twitter tweet </p></a></blockquote><h2 class="heading" style="text-align:left;" id="an-openweights-mo-e-trained-on-text">An open-weights MoE trained on text, images, audio, and video</h2><p class="paragraph" style="text-align:left;">Inkling is a Mixture-of-Experts transformer with 256 routed experts and two shared experts per layer, activating six routed experts per token. It was pretrained on 45 trillion tokens spanning text, images, audio, and video, then post-trained with large-scale reinforcement learning across more than 30 million rollouts, according to the <a class="link" href="https://thinkingmachines.ai/news/introducing-inkling/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned" target="_blank" rel="noopener noreferrer nofollow">model announcement</a>. Training ran on NVIDIA GB300 NVL72 systems.</p><p class="paragraph" style="text-align:left;">The model handles a context window of up to 1 million tokens and accepts multimodal input: images as 40x40 pixel patches, audio as spectrograms, and video. It also exposes a controllable &quot;thinking effort&quot; setting that lets users trade off how many tokens the model spends reasoning before it answers.</p><p class="paragraph" style="text-align:left;">Inkling joins a run of open-weight releases aimed at teams that want to run and modify models on their own hardware. We covered <a class="link" href="https://www.vp-land.com/p/google-s-gemma-4-might-be-the-most-capable-model-you-can-run-locally?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned" target="_blank" rel="noopener noreferrer nofollow">Google&#39;s Gemma 4</a>, which shipped under an Apache 2.0 license for local use without per-token fees.</p><h2 class="heading" style="text-align:left;" id="strong-reasoning-and-coding-scores-">Strong reasoning and coding scores, from a company that says it isn&#39;t the best</h2><p class="paragraph" style="text-align:left;">On benchmarks run at a high thinking-effort setting, Inkling reports 97.1% on AIME 2026, 87.2% on GPQA Diamond, and 77.6% on SWEBench Verified for agentic coding. With tool use, it scores 46.0% on Humanity&#39;s Last Exam. On multimodal tests it reports 73.5% on MMMU Pro for vision and 91.4% on VoiceBench for audio.</p><p class="paragraph" style="text-align:left;">Thinking Machines is candid about where the model sits, stating that Inkling is &quot;not the strongest overall model available today, open or closed.&quot; The pitch behind that framing, <a class="link" href="https://techcrunch.com/2026/07/15/thinking-machines-amps-up-its-bet-against-one-size-fits-all-ai-with-its-first-open-model-inkling/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned" target="_blank" rel="noopener noreferrer nofollow">as TechCrunch reported</a>, is a bet that &quot;AI that organizations can adapt for themselves will outperform the one-size-fits-all models the biggest labs currently sell.&quot;</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="access-runs-through-tinker-hosted-a">Access runs through Tinker, hosted APIs, and a free Playground</h2><p class="paragraph" style="text-align:left;">Teams can fine-tune Inkling on Tinker, which Thinking Machines is offering at a 50% limited-time discount with 64K and 256K context-length options. TechCrunch noted that Tinker, not the model itself, is where the company&#39;s <a class="link" href="https://techcrunch.com/2026/07/15/thinking-machines-amps-up-its-bet-against-one-size-fits-all-ai-with-its-first-open-model-inkling/?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned" target="_blank" rel="noopener noreferrer nofollow">revenue has to come from</a>, through training, fine-tuning, and a cut of the hosting ecosystem around the weights.</p><p class="paragraph" style="text-align:left;">For inference, the weights are hosted on TogetherAI, Fireworks, Modal, Databricks, and Baseten, and run on open tooling including SGLang, vLLM, and llama.cpp. A free Inkling Playground offers a chat interface with web search for a limited time.</p><p class="paragraph" style="text-align:left;">Thinking Machines Lab, led by former OpenAI CTO Mira Murati, built Tinker as its first product before releasing any model of its own. Inkling is its first public model.</p><h2 class="heading" style="text-align:left;" id="what-inkling-gives-creativetech-tea">What Inkling gives creative-tech teams</h2><p class="paragraph" style="text-align:left;">For studios and product teams weighing where to run AI, an open-weights multimodal model with downloadable weights changes the calculation. Teams can fine-tune Inkling on their own data, run it on their own hardware without per-token API fees, and point one base at text, image, audio, and video tasks.</p><p class="paragraph" style="text-align:left;">The constraint is scale. A 975-billion-parameter model needs serious compute to serve, and the smaller Inkling-Small preview points to where more accessible versions may land. For teams already working with <a class="link" href="https://www.vp-land.com/p/ideogram-4-0-ships-with-downloadable-weights-across-every-plan-and-the-api?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=thinking-machines-releases-inkling-its-first-open-weights-model-built-to-be-fine-tuned" target="_blank" rel="noopener noreferrer nofollow">downloadable model weights</a>, Inkling adds a multimodal reasoning option with a permissive release and a customization platform built around it.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Foundry Launches Griptape Enterprise, Putting Governed AI Pipelines Inside the Studio Firewall</title>
  <description></description>
      <enclosure length="344308" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/73f15fe3-8dc8-4cf8-aa79-d2a8a7ac25ce/unnamed__12_.jpg"/>
  <link>https://newsletter.vp-land.com/p/foundry-launches-griptape-enterprise-putting-governed-ai-pipelines-inside-the-studio-firewall</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/foundry-launches-griptape-enterprise-putting-governed-ai-pipelines-inside-the-studio-firewall</guid>
  <pubDate>Fri, 17 Jul 2026 05:31:32 +0000</pubDate>
  <atom:published>2026-07-17T05:31:32Z</atom:published>
    <category><![CDATA[Vfx Industry]]></category>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Foundry launched Griptape Enterprise, a new license tier of its Griptape AI workflow orchestration platform built to run generative pipelines entirely inside a studio&#39;s own infrastructure. The tier is available now.</p><ul><li><p class="paragraph" style="text-align:left;"><b>On-premises or private-cloud deployment</b> keeps plates, models, and metadata behind the studio firewall.</p></li><li><p class="paragraph" style="text-align:left;"><b>Modular diffusion pipelines</b> run FLUX, Stable Diffusion XL, LTX, and WAN models locally on studio GPUs.</p></li><li><p class="paragraph" style="text-align:left;"><b>Kyle Roche, Griptape co-founder,</b> steps into a new role as Foundry&#39;s Chief AI Officer.</p></li></ul><p class="paragraph" style="text-align:left;">Foundry positions it, in its <a class="link" href="https://www.foundry.com/news-and-awards/foundry-announces-griptape-enterprise?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-launches-griptape-enterprise-putting-governed-ai-pipelines-inside-the-studio-firewall" target="_blank" rel="noopener noreferrer nofollow">launch announcement</a>, as a governed alternative to the improvised setups studios have been building to bolt AI models onto existing pipelines.</p><h2 class="heading" style="text-align:left;" id="a-studiocontrolled-license-tier-aim">A studio-controlled license tier aimed at security and governance</h2><p class="paragraph" style="text-align:left;">Griptape Enterprise sits on top of the base Griptape platform and adds the controls high-end facilities need before they let generative tools touch client footage. It offers studio-controlled licensing and permissions, plus enterprise governance with centralized, unlimited licensing across a facility.</p><p class="paragraph" style="text-align:left;">Deployment runs on-premises or in a private cloud, so proprietary plates and trained models never leave studio-owned hardware. Color handling uses OpenColorIO for studio-grade plate and color management, and show-ready projects ship with templates for workspaces and versioning.</p><p class="paragraph" style="text-align:left;">We covered <a class="link" href="https://www.vp-land.com/p/foundry-acquires-griptape-to-accelerate-ai-integration-in-vfx-and-animation-pipelines?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-launches-griptape-enterprise-putting-governed-ai-pipelines-inside-the-studio-firewall" target="_blank" rel="noopener noreferrer nofollow">Foundry&#39;s acquisition of Griptape</a>, the deal that brought the orchestration technology in-house and set up this enterprise tier.</p><h2 class="heading" style="text-align:left;" id="nuke-integration-publishes-workflow">Nuke integration publishes workflows as versioned gizmos</h2><p class="paragraph" style="text-align:left;">Griptape Enterprise connects directly to Nuke. Artists can publish workflows as versioned gizmos and run scripts headlessly, which pushes generative steps into the same versioned, repeatable structure compositors already use for the rest of a shot.</p><p class="paragraph" style="text-align:left;">The tier also adds agentic automation with a choice of local or cloud LLMs, and it ships an MCP server that lets AI agents pilot Griptape at scale. We previously detailed how Foundry brought <a class="link" href="https://www.vp-land.com/p/foundry-brings-griptape-ai-agents-into-nuke-blender-and-maya-via-mcp?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-launches-griptape-enterprise-putting-governed-ai-pipelines-inside-the-studio-firewall" target="_blank" rel="noopener noreferrer nofollow">Griptape agents into Nuke</a>, Blender, and Maya through MCP, the same protocol underneath this release.</p><h2 class="heading" style="text-align:left;" id="modular-diffusion-pipelines-run-ope">Modular diffusion pipelines run open models on studio GPUs</h2><p class="paragraph" style="text-align:left;">The generative core is a set of modular diffusion pipelines that support FLUX, Stable Diffusion XL, LTX, and WAN models running locally on studio GPUs. Keeping inference on local hardware matters for facilities that cannot send frames to an outside API for legal, security, or client-confidentiality reasons.</p><p class="paragraph" style="text-align:left;">That local-first design lines up with Foundry&#39;s broader push to fold machine learning into its compositor, which we saw in <a class="link" href="https://www.vp-land.com/p/foundry-releases-nuke-17-0-with-native-gaussian-splat-support-and-new-usd-based-3d-system?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-launches-griptape-enterprise-putting-governed-ai-pipelines-inside-the-studio-firewall" target="_blank" rel="noopener noreferrer nofollow">Nuke 17.0</a> and its expanded machine-learning toolset.</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="kyle-roche-moves-from-griptape-ceo-">Kyle Roche moves from Griptape CEO to Foundry Chief AI Officer</h2><p class="paragraph" style="text-align:left;">Kyle Roche, Griptape co-founder and its former CEO, has been named Foundry&#39;s Chief AI Officer. He joined the company in February 2026 when the acquisition completed.</p><p class="paragraph" style="text-align:left;">Roche framed the enterprise tier as a way to move studios into AI production without the one-off workarounds many have assembled on their own:</p><div class="blockquote"><blockquote class="blockquote__quote"><p class="paragraph" style="text-align:left;">Foundry has been a trusted technology partner to visual effects and animation studios for over 30 years. With the new Griptape Enterprise license tier, Foundry is shepherding high-end creative teams into AI-enhanced production in a safe and secure way, with artistry and craft at its core. While studios are navigating how best to integrate AI models with bespoke platforms and workarounds, we&#39;re pleased to offer Griptape as a highly customizable and safe off-the-shelf alternative.</p><figcaption class="blockquote__byline"> Kyle Roche, Co-founder and Chief AI Officer </figcaption></blockquote></div><h2 class="heading" style="text-align:left;" id="what-governed-offtheshelf-ai-means-">What governed, off-the-shelf AI means for studio pipelines</h2><p class="paragraph" style="text-align:left;">For facilities weighing generative tools, the constraint has rarely been model quality. It has been control: where the data lives, who holds the license, and whether an AI step can slot into a versioned pipeline without breaking review and audit trails. Griptape Enterprise answers those questions by keeping deployment, models, and color management inside the studio&#39;s own walls, and by wiring the whole thing into Nuke&#39;s existing gizmo and scripting workflow.</p><p class="paragraph" style="text-align:left;">Whether high-end studios adopt a packaged orchestration layer over their own internal tooling is the open question. With FLUX, Stable Diffusion XL, LTX, and WAN all running locally and Nuke as the front end, Foundry is betting that the facilities most cautious about AI are the ones most likely to pay for a governed way in.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Foundry Ships SmartRoto for Nuke, Promising Up to 4x Faster Rotoscoping</title>
  <description></description>
      <enclosure length="257522" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0c8f5a0a-0cdf-4c0b-b010-5f309ba1913f/unnamed__11_.jpg"/>
  <link>https://newsletter.vp-land.com/p/foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping</guid>
  <pubDate>Fri, 17 Jul 2026 05:27:29 +0000</pubDate>
  <atom:published>2026-07-17T05:27:29Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Foundry has released SmartRoto, an AI rotoscoping add-on for Nuke that fills in the frames between an artist&#39;s keyframes. According to <a class="link" href="https://www.foundry.com/insights/machine-learning/smartroto-enabling-rotoscoping?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping" target="_blank" rel="noopener noreferrer nofollow">Foundry</a>, the tool moves and deforms splines across a shot from a few input frames, while leaving the artist in control of the final shape.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Up to 4x faster.</b> Foundry says SmartRoto can cut rotoscoping time to roughly a quarter of a manual pass.</p></li><li><p class="paragraph" style="text-align:left;"><b>Runs locally, trained on licensed data.</b> The model processes shots on the artist&#39;s own machine, and Foundry says it was trained on licensed footage.</p></li><li><p class="paragraph" style="text-align:left;"><b>$499 to start.</b> The introductory licence is priced at $499 and rises to $599, with a 90-day trial that closes August 16, 2026.</p></li></ul><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/6ec6B6UAtWM" width="100%"></iframe><h2 class="heading" style="text-align:left;" id="how-smart-roto-handles-the-inbetwee">How SmartRoto handles the in-between frames</h2><p class="paragraph" style="text-align:left;">Rotoscoping is the manual work of drawing and animating mattes around objects so compositors can isolate them, and it remains one of the most time-intensive tasks in a VFX pipeline. SmartRoto targets the slowest part of that job: the tweening between an artist&#39;s keyframes.</p><p class="paragraph" style="text-align:left;">The artist still sets up the initial splines and a small number of keyframes. SmartRoto then predicts the intermediate shapes across the sequence, tracking and deforming the splines to follow motion in the plate. Foundry says the artist keeps control throughout and can correct any frame, so the tool assists the roto pass rather than replacing the artist&#39;s judgment on tricky edges.</p><h2 class="heading" style="text-align:left;" id="from-a-dneg-research-project-to-a-s">From a DNEG research project to a shipping product</h2><p class="paragraph" style="text-align:left;">SmartRoto did not start as a product. It began as a Foundry research collaboration with visual effects studio DNEG and the University of Bath, documented in Foundry&#39;s own <a class="link" href="https://www.foundry.com/insights/machine-learning/smartroto-enabling-rotoscoping?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping" target="_blank" rel="noopener noreferrer nofollow">research write-up</a>. The model was trained on DNEG&#39;s production rotoscopy artwork, a dataset the team describes as more than 650,000 artist-animated shapes and 125 million keyframes drawn from real shot work.</p><p class="paragraph" style="text-align:left;">The research goal was more modest than the launch claim. Ben Kent, a research engineering manager at Foundry, framed the original target as saving &quot;even 25% of an artist&#39;s time.&quot; The commercial release now advertises up to a 4x speedup, a larger figure that VFX teams will want to test against their own footage before budgeting around it.</p><h2 class="heading" style="text-align:left;" id="part-of-foundrys-longer-ai-buildout">Part of Foundry&#39;s longer AI build-out in Nuke</h2><p class="paragraph" style="text-align:left;">SmartRoto lands on top of years of machine learning work inside Nuke. Foundry&#39;s AI lineage in the compositor traces back to CopyCat, and the company has kept expanding those capabilities. We covered <a class="link" href="https://www.vp-land.com/p/foundry-releases-nuke-17-0-with-native-gaussian-splat-support-and-new-usd-based-3d-system?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping" target="_blank" rel="noopener noreferrer nofollow">Nuke 17.0</a>, which added native Gaussian Splat support and broadened its machine learning toolset.</p><p class="paragraph" style="text-align:left;">The company has also been buying and building around AI infrastructure. We reported on Foundry&#39;s <a class="link" href="https://www.vp-land.com/p/foundry-acquires-griptape-to-accelerate-ai-integration-in-vfx-and-animation-pipelines?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping" target="_blank" rel="noopener noreferrer nofollow">Griptape acquisition</a>, which brought Python-based agent orchestration into its pipeline tools. SmartRoto fits the same pattern: targeted AI applied to a specific, repetitive artist task rather than a general-purpose generator.</p><h2 class="heading" style="text-align:left;" id="ai-roto-joins-a-crowded-problem-spa">AI roto joins a crowded problem space</h2><p class="paragraph" style="text-align:left;">SmartRoto enters a category where other teams are also attacking hard matte work. We previously covered <a class="link" href="https://www.vp-land.com/p/corridor-crew-open-sources-corridorkey-a-neural-network-tool-that-solves-green-screen-s-hardest-prob?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=foundry-ships-smartroto-for-nuke-promising-up-to-4x-faster-rotoscoping" target="_blank" rel="noopener noreferrer nofollow">CorridorKey</a>, the open-source neural keyer that tackles green screen edges many traditional keyers struggle with. Where CorridorKey focuses on keying, SmartRoto focuses on the spline-based roto that artists fall back on when keying is not an option.</p><p class="paragraph" style="text-align:left;">For studios weighing it, the practical questions are cost and fit. At an introductory $499 per licence and a local install that keeps footage off external servers, SmartRoto is priced for individual artists and small teams as much as large facilities. The 90-day trial gives roto artists a window to run it against their own shots before the price steps up to $599, and before the August 16, 2026 deadline closes.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>We Tested Seedream 5.0 Pro While Meta Backpedaled on Instagram AI Remixes</title>
  <description></description>
      <enclosure length="1659030" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0b15ccf6-e4ab-4ace-8947-77cbf3178e21/comfy_mcp.jpg"/>
  <link>https://newsletter.vp-land.com/p/we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes</guid>
  <pubDate>Thu, 16 Jul 2026 16:46:48 +0000</pubDate>
  <atom:published>2026-07-16T16:46:48Z</atom:published>
    <category><![CDATA[Denoised Podcast]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">This week runs from the practical to the existential: a new image model that edits from annotations, Meta walking back an opt-out AI feature after public backlash, MCP arriving inside Unreal Engine and ComfyUI, and a research claim about a &quot;hidden brain&quot; that has AGI watchers talking.</p><p class="paragraph" style="text-align:left;">We test <b>Seedream 5.0 Pro&#39;s</b> precise editing live, break down <b>Meta Muse</b> and the Instagram remix reversal, walk through what <b>MCP</b> actually changes for Unreal and ComfyUI workflows, and dig into the <b>J-Space</b> claim about a reasoning structure that appeared inside a neural network on its own.</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/pmuDx9mo-Tw" width="100%"></iframe><div class="section" style="background-color:transparent;margin:30.0px 30.0px 30.0px 30.0px;padding:0.0px 0.0px 0.0px 0.0px;"><table width="100%" class="bh__column_wrapper"><tr><td width="33%" class="bh__column"><div class="image"><a class="image__link" href="https://vpgo.link/denoised-spotify?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/73810210-52ee-4c5a-bfc9-ecfbf1e97f7a/Spotify_-_04.png?t=1739592377"/></a></div></td><td width="33%" class="bh__column"><div class="image"><a class="image__link" href="https://vpgo.link/denoised-apple?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5cf3881d-77c4-4364-a1d6-679d33cf5734/Apple_-_04.png?t=1739592405"/></a></div></td><td width="33%" class="bh__column"><div class="image"><a class="image__link" href="https://vpgo.link/denoised-youtube?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ce4801d4-82ec-413c-bdb9-24813d091e59/YouTube_-_04.png?t=1739592417"/></a></div></td></tr></table></div><h2 class="heading" style="text-align:left;" id="quick-take">Quick Take</h2><p class="paragraph" style="text-align:left;">Four stories, one throughline: the tools keep closing the gap on precision while the questions around consent, cost, and control get harder. We run a new image model against a period-accuracy test, weigh an opt-out remix feature that pulled in a SAG statement, separate where agentic tool access beats a plain API, and sit with a research claim that is either a real step toward machine reasoning or an investor talking point. The capabilities are arriving faster than the frameworks for using them.</p><h2 class="heading" style="text-align:left;" id="what-we-tested-seedream-50-pros-ann">What We Tested: Seedream 5.0 Pro&#39;s Annotation Edits</h2><p class="paragraph" style="text-align:left;">ByteDance shipped <b>Seedream 5.0 Pro</b>, and the headline for Addy is precision. You can target a specific change in an image and the model applies it without disturbing the rest of the frame.</p><p class="paragraph" style="text-align:left;">The part worth noting is how it takes direction. It does not need a mask. Mark up an image with a red box around a region, then write a prompt like &quot;change the red box to something else,&quot; and the model reads the annotation and follows it.</p><p class="paragraph" style="text-align:left;"><b>The capability jump.</b> Joey put that annotation control in the same league as GPT and Nano Banana Pro, and Addy argued it lands even more precise, with better text rendering and stronger photorealism on text-to-image than earlier versions. When we <a class="link" href="https://www.vp-land.com/p/is-seedream-the-next-nano-banana-plus-comfy-cloud-and-more-film-tech-updates?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">covered Seedream 4</a>, Nano Banana beat it in a head-to-head; Addy&#39;s read is that the gap has since closed on targeted edits.</p><p class="paragraph" style="text-align:left;"><b>The 1970s New York test.</b> Joey ran his standard litmus prompt live: a busy 1970s New York City street with taxi cabs and pedestrians. The result captured the period vibe, put a recognizable skyline down the middle, and kept modern buildings out of frame. It also showed the model&#39;s weak spots.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Background softness.</b> Fine detail stayed fuzzy, a trait Addy has flagged since <a class="link" href="https://www.vp-land.com/p/bytedance-s-seedream-4-5-levels-up-text-and-quality?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Seedream 4.5</a>, and small signage text came out blurred.</p></li><li><p class="paragraph" style="text-align:left;"><b>Physical logic.</b> The cars pointed in directions that did not make sense, with a near-collision in the middle of the frame.</p></li><li><p class="paragraph" style="text-align:left;"><b>A cultural gap.</b> Both of us landed on the same observation: American models render the gritty, high-contrast Taxi Driver version of 1970s New York, while the output here read cleaner and more idyllic. Joey&#39;s theory is a training-data nuance, since a Chinese model trains on less American reference material.</p></li></ul><p class="paragraph" style="text-align:left;"><b>Resolution ceiling.</b> Seedream 5.0 Pro tops out around 2K and needs an upscale to reach 4K, where <a class="link" href="https://www.vp-land.com/p/nano-banana-pro-meta-s-sam-3d-and-world-labs-marble?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Nano Banana Pro</a> renders 4K natively. Addy expects native high resolution to follow, since ByteDance&#39;s video side already does 4K.</p><p class="paragraph" style="text-align:left;"><b>The workflow pairing.</b> The natural use is to generate keyframes in Seedream and move them into Seedance as start and end frames. Looking ahead, Addy pointed to rumors he called &quot;pretty confirmed&quot; that a coming ByteDance video model will accept up to 50 reference images and generate clips as long as three minutes. Treat that as unconfirmed for now. His caveat is real either way: the model still tries to use every reference you feed it rather than choosing the right ones per shot, and temporal control over which image lands at which second is still missing.</p><h2 class="heading" style="text-align:left;" id="what-we-debated-meta-muse-and-the-i">What We Debated: Meta Muse and the Instagram Remix Reversal</h2><p class="paragraph" style="text-align:left;">Meta is back on the board with <b>Muse</b>, a new image and video model out of its Superintelligence Labs. Meta describes it as agentic image generation, which means the model calls tools based on the prompt rather than generating in one pass. The showcased examples include sharp text rendering and generating a working QR code inside an image. We <a class="link" href="https://www.vp-land.com/p/meta-launches-agentic-muse-image-and-previews-muse-video-from-superintelligence-labs?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">covered the Muse Image launch</a> when it dropped.</p><p class="paragraph" style="text-align:left;">On raw output, Addy&#39;s early read is that Muse looks on par with what is already available. The significance is positioning: after a stretch of falling behind, Meta has a competitive model again.</p><p class="paragraph" style="text-align:left;"><b>Where the business logic points.</b> Joey connected Muse to reports that Meta is monetizing spare data center capacity, an AWS-style move that suggests its own internal AI demand has slowed. The counterweight is Meta&#39;s advertising machine. A native model wired into its ad tools fits the company&#39;s MarTech strength, and both of us landed on the target user being small advertisers.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Not the enterprise buyer.</b> Coca-Cola-scale brands run in-house AI teams and will not need this.</p></li><li><p class="paragraph" style="text-align:left;"><b>The micro and nano brands.</b> Mom-and-pop shops and small Etsy-style sellers already pay for Facebook and Instagram ads and lack creative capacity. Muse is a way for them to scale content, mirroring what Google offers advertisers.</p></li></ul><p class="paragraph" style="text-align:left;"><b>The remix controversy.</b> The friction point is a feature that lets people remix anyone else&#39;s Instagram photos with AI, launched as opt-out, meaning every account was included by default. After we recorded, Meta walked it back following online backlash, so accounts are no longer auto-enrolled. SAG issued a statement, since the default exposed any actor or public figure on the platform.</p><p class="paragraph" style="text-align:left;"><b>Joey&#39;s concern:</b> the harm does not require an extreme edit. A teenager&#39;s photos could be remixed into something merely embarrassing, which current safeguards built for clearly harmful content would not catch.</p><p class="paragraph" style="text-align:left;"><b>Addy&#39;s counter:</b> anyone could already screenshot an Instagram photo and run it through another model. The opt-out design mostly removed friction rather than creating a new capability.</p><p class="paragraph" style="text-align:left;">That tension pulled us to a broader number Joey raised: a reported figure that roughly <b>20% of YouTube is already AI content</b>. Addy suspects that undercounts, and both of us noted the open question is not how much AI gets uploaded but how much gets watched, with a likely swing back toward authentic, personality-driven content.</p><div class="custom_html"><iframe src="https://embeds.beehiiv.com/9016a043-cf4f-4c0f-be10-46ce481e3460" data-test-id="beehiiv-embed" width="100%" height="320" frameborder="0" style="border-radius: 4px; border: 2px solid #e5e7eb; margin: 0; background-color: transparent;"></iframe></div><h2 class="heading" style="text-align:left;" id="what-we-explored-mcp-lands-in-unrea">What We Explored: MCP Lands in Unreal Engine and ComfyUI</h2><p class="paragraph" style="text-align:left;"><b>Model Context Protocol</b>, the standard Anthropic introduced in early 2025, moved from concept to production tooling this week, with beta MCP support arriving in <b>Unreal Engine</b> and MCP access now in <b>ComfyUI</b>.</p><p class="paragraph" style="text-align:left;">Addy&#39;s working definition of why it matters: a traditional API needs rigid, structured scripting to call anything, while MCP lets an agent ask a tool what it can do and act more fluidly, which fits how agents operate. In Unreal&#39;s beta, you can tell Claude to build a blueprint that does a specific thing and it builds it. In ComfyUI, you can describe a workflow and it assembles the nodes.</p><p class="paragraph" style="text-align:left;"><b>The ComfyUI caveat.</b> The integration currently works only with ComfyUI Cloud, not the desktop build. You can still build a workflow in the cloud and download it locally, which pairs with the direction we noted around <a class="link" href="https://www.vp-land.com/p/is-seedream-the-next-nano-banana-plus-comfy-cloud-and-more-film-tech-updates?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Comfy Cloud</a> earlier.</p><p class="paragraph" style="text-align:left;"><b>Where it helps most.</b> Joey&#39;s take is that MCP flattens the learning curve. It took him roughly three years to reach a generalist level in Unreal; he thinks the same competence is now reachable in closer to three months with the hours still required. For ComfyUI, the value is exploratory: most users understand a fraction of the available nodes, and an agent that knows every node can suggest which ones solve a given problem.</p><p class="paragraph" style="text-align:left;"><b>APIs are not going away.</b> Addy&#39;s caution: MCP calls burn tokens, so for predictable, repetitive work against structured data, a script or API is faster, more reliable, and cheaper. Anyone building for scale and customers still needs that reliability.</p><p class="paragraph" style="text-align:left;"><b>The bonus use case.</b> Addy described pointing Claude Code at a DaVinci Resolve project that kept crashing on open. Instead of the effect everyone suspected, the logs showed a clip that had gone offline while Resolve still treated it as online, and the fix was to mark the clip offline. A half-day support headache turned into a few minutes of log reading.</p><h2 class="heading" style="text-align:left;" id="what-we-questioned-the-j-space-agi-">What We Questioned: The &quot;J-Space&quot; AGI Claim</h2><p class="paragraph" style="text-align:left;">The episode&#39;s most speculative thread is <b>J-Space</b>. Joey described a claimed Anthropic research finding: researchers examining a neural network reportedly located a hidden layer, a reasoning structure that developed inside the model on its own and was not part of the original design. Delete it and the model reportedly loses the ability to reason; restore it and reasoning returns.</p><p class="paragraph" style="text-align:left;">Treat this as unverified. Joey flagged it directly as something that could be an investor talking point, and neither of us had fully read the paper on air. Addy also recalled a related claim that the model&#39;s internal reasoning ran in a compressed shorthand of English that used fewer tokens.</p><p class="paragraph" style="text-align:left;"><b>The question:</b> if a reasoning structure really can emerge on its own, it is a signal AGI watchers point to. If it cannot be reproduced, it is a headline. Either way, the image and video tools filmmakers actually use keep advancing regardless of where the AGI debate lands.</p><p class="paragraph" style="text-align:left;">We closed on a lighter forward-looker: The Verge reported that <b>Character AI</b> is moving into interactive microdramas, a game-like format you can steer rather than just watch, which lines up with the commercial microdrama use cases we keep circling back to.</p><h2 class="heading" style="text-align:left;" id="bottom-line-capability-is-outrunnin">Bottom Line: Capability Is Outrunning the Guardrails</h2><p class="paragraph" style="text-align:left;">Every story this week is a version of the same gap between what these tools can do and the frameworks for doing it responsibly.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Seedream 5.0 Pro</b> proves annotation-based editing is now precise enough to compete, while the resolution ceiling and cultural blind spots show where a model&#39;s training still shows through.</p></li><li><p class="paragraph" style="text-align:left;"><b>Meta Muse</b> put a competitive model back in Meta&#39;s hands, but the opt-out remix rollout showed the company reaching for engagement ahead of consent, and the reversal came only after the backlash.</p></li><li><p class="paragraph" style="text-align:left;"><b>MCP</b> makes complex software approachable inside Unreal and ComfyUI, without retiring the APIs that still win on scale and cost.</p></li><li><p class="paragraph" style="text-align:left;"><b>J-Space</b> is either an early marker of machine reasoning or an investment narrative, and the claim is unverified until the research holds up.</p></li></ul><p class="paragraph" style="text-align:left;">The tools are getting easier and sharper. The decisions about consent, cost, and trust are getting harder.</p><h2 class="heading" style="text-align:left;" id="links-from-this-episode">Links from This Episode</h2><p class="paragraph" style="text-align:left;"><b>Tools & Platforms:</b></p><ul><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/bytedance-s-seedream-4-5-levels-up-text-and-quality?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Seedream 4.5</a>. VP Land coverage of the prior ByteDance image release.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/meta-launches-agentic-muse-image-and-previews-muse-video-from-superintelligence-labs?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Meta Muse Image and Muse Video</a>. VP Land coverage of the Superintelligence Labs launch.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/seedance-2-0?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Seedance 2.0</a>. VP Land coverage of ByteDance&#39;s video model.</p></li></ul><p class="paragraph" style="text-align:left;"><b>Models & Comparisons:</b></p><ul><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/nano-banana-pro-meta-s-sam-3d-and-world-labs-marble?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Nano Banana Pro</a>. VP Land coverage of Google&#39;s native-4K image model.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.vp-land.com/p/is-seedream-the-next-nano-banana-plus-comfy-cloud-and-more-film-tech-updates?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=we-tested-seedream-5-0-pro-while-meta-backpedaled-on-instagram-ai-remixes" target="_blank" rel="noopener noreferrer nofollow">Is Seedream the Next Nano Banana? Plus Comfy Cloud</a>. Our earlier Seedream 4 test and Comfy Cloud context.</p><p class="paragraph" style="text-align:left;"></p></li></ul></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

      <item>
  <title>Mirelo and Kyutai Open-Source MuScriptor, an Audio-to-MIDI Model That Transcribes Full Mixes</title>
  <description></description>
      <enclosure length="338883" type="image/jpeg" url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/4fec4d6b-d96d-469e-bfd9-01e9576be36a/unnamed__10_.jpg"/>
  <link>https://newsletter.vp-land.com/p/mirelo-and-kyutai-open-source-muscriptor-an-audio-to-midi-model-that-transcribes-full-mixes</link>
  <guid isPermaLink="true">https://newsletter.vp-land.com/p/mirelo-and-kyutai-open-source-muscriptor-an-audio-to-midi-model-that-transcribes-full-mixes</guid>
  <pubDate>Tue, 14 Jul 2026 15:14:56 +0000</pubDate>
  <atom:published>2026-07-14T15:14:56Z</atom:published>
    <category><![CDATA[Article]]></category>
    <category><![CDATA[Generative Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Mirelo AI and Kyutai <a class="link" href="https://www.mirelo.ai/blog/turning-audio-to-midi?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=mirelo-and-kyutai-open-source-muscriptor-an-audio-to-midi-model-that-transcribes-full-mixes" target="_blank" rel="noopener noreferrer nofollow">open-sourced MuScriptor</a>, an audio-to-MIDI model that takes a finished music mix and transcribes every instrument at once into separate MIDI tracks. Most audio-to-MIDI tools handle one sound at a time. MuScriptor reads the whole mix.</p><p class="paragraph" style="text-align:left;"><b>The model outputs a separate MIDI track per instrument on a piano roll, with automatic instrument detection and labeling.</b> It also detects chords, key, and tempo, and supports one-click export. The weights are on Hugging Face, the inference code is on GitHub, and a companion tool ships free inside Mirelo Studio.</p><h2 class="heading" style="text-align:left;" id="transcription-treated-as-a-language">Transcription treated as a language modeling problem</h2><p class="paragraph" style="text-align:left;">MuScriptor uses a decoder-only transformer that accepts mel-spectrograms and autoregressively generates token sequences representing pitch, timing, and instrument type. Mirelo frames the approach as &quot;music transcription as language modeling,&quot; borrowing the same architecture pattern that powers text generation and applying it to notes instead of words.</p><p class="paragraph" style="text-align:left;">The design targets a specific failure point. Dense, multi-instrument audio has historically tripped up transcription models built to isolate a single voice or instrument. According to Mirelo, MuScriptor &quot;takes the full mix and transcribes every instrument at once&quot; rather than requiring source separation as a first step.</p><h2 class="heading" style="text-align:left;" id="a-threestage-training-pipeline-mixi">A three-stage training pipeline mixing synthetic and real audio</h2><p class="paragraph" style="text-align:left;">The model was trained across three datasets, each doing a different job:</p><ul><li><p class="paragraph" style="text-align:left;"><b>DSynth.</b> Roughly 1.45 million MIDI files rendered to audio with on-the-fly synthesis, giving the model broad coverage of note and instrument combinations.</p></li><li><p class="paragraph" style="text-align:left;"><b>DReal.</b> About 170,000 real recordings, more than 11,000 hours, paired with aligned annotations so the model learns how studio audio actually sounds.</p></li><li><p class="paragraph" style="text-align:left;"><b>DRL.</b> A smaller set of 300 high-quality tracks used for a GRPO-like reinforcement learning stage to sharpen output quality.</p></li></ul><p class="paragraph" style="text-align:left;">Fine-tuning on real recordings mattered. Mirelo reports that training on real data improved transcription metrics by roughly 20 percentage points over a version pre-trained on synthetic audio alone, per the team&#39;s <a class="link" href="https://arxiv.org/abs/2607.08168?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=mirelo-and-kyutai-open-source-muscriptor-an-audio-to-midi-model-that-transcribes-full-mixes" target="_blank" rel="noopener noreferrer nofollow">ArXiv paper</a>, which credits IRCAM in its citation.</p><h2 class="heading" style="text-align:left;" id="why-perinstrument-midi-matters-for-">Why per-instrument MIDI matters for scoring and remixing</h2><p class="paragraph" style="text-align:left;">For composers, editors, and music supervisors, MIDI is editable in a way that audio is not. A transcription that separates a mix into per-instrument tracks turns a reference recording into a starting point: swap a synth patch, re-voice a chord, quantize a drum pattern, or lift a bassline into a new arrangement. Getting there previously meant either manual transcription or stems that still had to be transcribed one instrument at a time.</p><p class="paragraph" style="text-align:left;">MuScriptor lands amid a run of open and accessible music AI. We covered <a class="link" href="https://www.vp-land.com/p/google-launches-lyria-3-gemini-s-high-fidelity-ai-music-generator?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=mirelo-and-kyutai-open-source-muscriptor-an-audio-to-midi-model-that-transcribes-full-mixes" target="_blank" rel="noopener noreferrer nofollow">Google&#39;s Lyria 3</a>, which generates full tracks with vocals from text prompts.</p><p class="paragraph" style="text-align:left;">We also covered <a class="link" href="https://www.vp-land.com/p/elevenlabs-lets-you-train-a-custom-ai-music-model-on-your-own-catalog?utm_source=newsletter.vp-land.com&utm_medium=newsletter&utm_campaign=mirelo-and-kyutai-open-source-muscriptor-an-audio-to-midi-model-that-transcribes-full-mixes" target="_blank" rel="noopener noreferrer nofollow">ElevenLabs Music Finetunes</a>, which trains a custom music model on a user&#39;s own catalog. MuScriptor works the opposite direction: instead of generating audio, it reads existing audio back into a symbolic, editable format.</p><h2 class="heading" style="text-align:left;" id="open-weights-lower-the-barrier-for-">Open weights lower the barrier for building on top</h2><p class="paragraph" style="text-align:left;">Because MuScriptor ships with open weights and inference code, developers can run it locally or wire it into their own tools rather than depending on a hosted API. The decoder-only, language-modeling framing also means the architecture will feel familiar to teams already working with transformer models.</p><p class="paragraph" style="text-align:left;">For anyone who wants to try it without touching code, the audio-to-MIDI tool in Mirelo Studio is free and runs the improved model. The combination of a free hosted tool and open weights gives both casual users and builders a way in.</p></div></div>
  ]]></content:encoded>
<author>podcast@newterritory.media (Coffee and Celluloid)</author><itunes:explicit>no</itunes:explicit><itunes:author>Coffee and Celluloid</itunes:author><itunes:keywords>film,photography,art,design,interview,filmmaking,independent</itunes:keywords></item>

  </channel>
</rss>