{
  "id": "creative-generative-media",
  "label": "Creative & Generative Media",
  "description": "AI for generating and editing images, video, audio, 3D assets, and cross-media content. Mostly leading-edge with rapid advancement — image generation, music composition, and voice synthesis are approaching good practice. Video generation and 3D asset creation are progressing fast but quality and controllability gaps persist. The most active domain by momentum: over half the practices are advancing.",
  "icon": "🎬",
  "filters": [
    "creating"
  ],
  "hasSummary": true,
  "hasExecSummary": true,
  "practiceCount": 21,
  "evidenceCount": 3809,
  "practices": [
    {
      "slug": "3d-asset-scene-and-texture-generation",
      "name": "3D asset, scene & texture generation",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that generates 3D models, scenes, and textures from text descriptions, images, or procedural rules. Includes text-to-3D pipelines and PBR material generation; distinct from game and AR/VR content which targets interactive rather than static 3D output.",
      "evidenceCount": 179
    },
    {
      "slug": "ai-driven-video-editing-and-post-production",
      "name": "AI-driven video editing & post-production",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that automates video editing including highlight detection, compilation, colour grading, transitions, and post-production workflows. Includes automated sports highlight reels and AI-assisted colour and audio correction; distinct from video generation which creates new footage rather than editing existing material.",
      "evidenceCount": 187
    },
    {
      "slug": "audio-production-editing-podcasts-and-sound-design",
      "name": "Audio production — editing, podcasts & sound design",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": "market-scope",
      "description": "AI that removes noise, enhances audio quality, automates podcast production workflows, and generates sound effects and designs. Includes automated mastering and AI-assisted foley; distinct from music generation which creates melodic and harmonic compositions.",
      "evidenceCount": 179
    },
    {
      "slug": "avatar-generation-and-personalised-media-at-scale",
      "name": "Avatar generation & personalised media at scale",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that creates virtual presenters, digital avatars, and personalised media variants at scale for marketing, training, and communication. Includes photorealistic talking-head generation and dynamic content personalisation; distinct from video generation which creates general rather than personalised or avatar-based content.",
      "evidenceCount": 178
    },
    {
      "slug": "brand-asset-generation-and-variation",
      "name": "Brand asset generation & variation",
      "tier": "good-practice",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of brand assets, template variations, and guideline-compliant creative materials across formats. Includes automated asset resizing and brand-consistent variant generation; distinct from brand-voice workflows which enforce written rather than visual brand standards.",
      "evidenceCount": 156
    },
    {
      "slug": "content-authenticity-deepfake-detection-and-provenance",
      "name": "Content authenticity — deepfake detection & provenance",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that detects deepfakes, authenticates content origin, and applies provenance metadata and watermarks to verify media integrity. Includes C2PA standard implementation and synthetic media detection; distinct from content safety which filters harmful outputs rather than verifying authenticity.",
      "evidenceCount": 213
    },
    {
      "slug": "image-editing-inpainting-outpainting-and-extension",
      "name": "Image editing — inpainting, outpainting & extension",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI-powered image editing for filling in, extending, and modifying images including background removal and replacement. Includes generative fill and canvas extension; distinct from style transfer which transforms the entire image rather than editing specific regions.",
      "evidenceCount": 176
    },
    {
      "slug": "image-editing-style-transfer-and-artistic-transformation",
      "name": "Image editing — style transfer & artistic transformation",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that transforms images between artistic styles, colour palettes, and visual treatments while preserving content. Includes neural style transfer and artistic filter application; distinct from inpainting which modifies content rather than visual style.",
      "evidenceCount": 199
    },
    {
      "slug": "image-generation-photorealistic-and-illustrative",
      "name": "Image generation — photorealistic & illustrative",
      "tier": "good-practice",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that generates photorealistic images and illustrations from text prompts or reference images. Includes diffusion-based generation for photography, concept art, and illustration styles; distinct from product visualisation which targets commercial product imagery.",
      "evidenceCount": 183
    },
    {
      "slug": "image-generation-product-visualisation-and-mockups",
      "name": "Image generation — product visualisation & mockups",
      "tier": "good-practice",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of product visualisations, mockups, and lifestyle imagery for e-commerce and marketing. Includes background replacement and lifestyle scene generation; distinct from virtual try-on which simulates wearing/using rather than displaying products.",
      "evidenceCount": 162
    },
    {
      "slug": "image-processing-upscaling-restoration-and-compositing",
      "name": "Image processing — upscaling, restoration & compositing",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that upscales low-resolution images, restores damaged photos, removes backgrounds, and composites elements. Includes super-resolution and intelligent matting; distinct from image editing which modifies creative content rather than performing technical processing.",
      "evidenceCount": 202
    },
    {
      "slug": "interactive-content-game-arvr-and-environment-generation",
      "name": "Interactive content — game, AR/VR & environment generation",
      "tier": "leading-edge",
      "trend": "slowing",
      "blockerType": null,
      "description": "AI that generates game assets, levels, AR/VR environments, and interactive content experiences. Includes procedural level generation and spatial AR content creation; distinct from 3D asset generation which creates individual objects rather than interactive experiences.",
      "evidenceCount": 176
    },
    {
      "slug": "lip-sync-and-video-dubbing",
      "name": "Lip sync & video dubbing",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that synchronises lip movements with dubbed audio for video localisation across languages. Includes face re-animation and multilingual dubbing; distinct from text-to-speech which generates audio without visual synchronisation.",
      "evidenceCount": 170
    },
    {
      "slug": "motion-capture-and-pose-estimation-for-content-production",
      "name": "Motion capture & pose estimation for content production",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI-powered markerless motion capture and pose estimation for animation, gaming, and video production. Includes single-camera mocap and real-time pose tracking; distinct from avatar generation which creates virtual characters rather than capturing real movement.",
      "evidenceCount": 209
    },
    {
      "slug": "multimodal-content-generation",
      "name": "Multimodal content generation",
      "tier": "good-practice",
      "trend": "accelerating",
      "blockerType": null,
      "description": "AI that generates integrated multi-format content combining text, images, and layout in a single workflow. Includes newsletter generation and social card creation; distinct from content repurposing which adapts existing content rather than generating multimodal output from scratch.",
      "evidenceCount": 178
    },
    {
      "slug": "music-generation-background-and-ambient",
      "name": "Music generation — background & ambient",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of background music, ambient soundscapes, and functional music for content and environments. Includes royalty-free generation and mood-based creation; distinct from full composition which produces standalone musical works.",
      "evidenceCount": 198
    },
    {
      "slug": "music-generation-full-composition",
      "name": "Music generation — full composition",
      "tier": "bleeding-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of complete musical compositions with melody, harmony, arrangement, and production. Includes genre-specific composition and multi-instrument arrangement; distinct from background music which produces functional rather than standalone works.",
      "evidenceCount": 171
    },
    {
      "slug": "text-to-speech-natural-voice-synthesis",
      "name": "Text-to-speech — natural voice synthesis",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of natural-sounding speech from text for audiobooks, accessibility, navigation, and content delivery. Includes multi-language synthesis and emotional expression; distinct from voice cloning which replicates specific voices rather than generating generic natural speech.",
      "evidenceCount": 215
    },
    {
      "slug": "text-to-speech-voice-cloning-and-custom-voices",
      "name": "Text-to-speech — voice cloning & custom voices",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI that clones specific voices or creates custom synthetic voices for branded content and personalisation. Includes few-shot voice cloning and brand voice creation; distinct from natural TTS which uses standard rather than replicated voices.",
      "evidenceCount": 165
    },
    {
      "slug": "video-generation-long-form-narrative-and-explainer",
      "name": "Video generation — long-form narrative & explainer",
      "tier": "bleeding-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of longer narrative videos, explainers, and educational content with coherent storylines. Includes multi-scene generation and narrative consistency; distinct from short-form which produces clips rather than structured narratives.",
      "evidenceCount": 159
    },
    {
      "slug": "video-generation-short-form",
      "name": "Video generation — short-form",
      "tier": "leading-edge",
      "trend": "steady",
      "blockerType": null,
      "description": "AI generation of short-form video content for social media, advertising, and promotional clips. Includes text-to-video and image-to-video generation for brief clips; distinct from long-form generation which produces extended narratives.",
      "evidenceCount": 154
    }
  ],
  "summary": "## Where AI Stands in Creative & Generative Media\n\nCreative and generative media is the domain where the capability question has been answered most decisively and the acceptance question least. In a blind test published in the NeurIPS Creative AI track, listeners could not reliably tell Suno output from human music. Voice clones reach 97% fidelity while humans detect only 37.5% of them. The leading text-to-video models sit in a statistical tie on independent arenas. Money has followed: Adobe reports AI-first annual recurring revenue above $650M, up 150% year on year; ElevenLabs says it is pacing at $600M and was valued at $22bn in a secondary share sale; Runway is at $200M, Kling reports $300M annualised and Suno $300M. Yet the same period produced a casualty list. OpenAI closed Sora, an app that earned $2.1M in total. Beatoven.ai shut its consumer music service in August. Soul Machines entered receivership in February. Adobe's share price fell from $630 in February 2024 to $266 by September 2026 even as its AI revenue grew.\n\nWhat works in production is bounded. WSC Sports generated more than 134,000 videos across the 104 matches of the 2026 World Cup; transcription and noise removal are dependable on clean audio; avatar video is replacing filming for corporate training and localisation; brands generate backgrounds, resizes and variants around a real photograph. Wherever output must be faithful to something, such as a product, a face, a brand colour, a storyline or a mesh that a rigger can use, a person still stands between generation and publication. Algorithmine finds 68% of enterprise video deployments need human intervention to reach publishable quality. Retailers manually correct 60-80% of inpainting outputs. Photoroom's 850-product benchmark found frontier editing models preserve product accuracy in 29% of outputs. Two areas sit outside that pattern. Workflows that produce text, image, audio and video from one brief are spreading fastest, because Adobe and Google have put them behind conversational assistants and into Google Ads and Workspace, although the UK's ONS still puts visual content creation at 16% of businesses with ten or more employees. Full-song generation is the reverse case: enormous consumer volume, almost no organisations reporting that it works in production, and 87% of Canadian musicians in a University of Alberta study viewing generative AI negatively.\n\nMomentum is building in distribution, not in models. Capability now arrives as a feature inside software people already pay for: five video models in Premiere's timeline, Lyria in Gemini, auto dubbing switched on by default at YouTube, markerless capture in Unreal Engine 5.8. That widens reach and squeezes specialists. The stall is on the demand side. Gartner finds half of consumers would rather buy from brands that keep AI out of public-facing content. Deezer says fully AI-generated tracks exceeded half of daily uploads on peak days in July and draw 1-3% of streams. Among GDC respondents, 52% now view generative AI negatively, up from 30% a year earlier, and Roblox has lost daily users for three consecutive quarters while expanding its AI creation tooling. What separates this domain from its neighbours is that its output is judged by audiences, not by the organisation that deployed it, and that it draws on voices, likenesses and catalogues belonging to someone else. Courts, collecting societies and regulators are rewriting the terms while the tools are already in use: EU AI Act transparency obligations have applied since 2 August 2026, and a Munich court ruled against Suno on 31 July.\n\n## What's New, 2026-09-17 to 2026-10-01\n\nThe fortnight's clearest pattern was consolidation around platform owners. Adobe closed its roughly $340M purchase of Topaz Labs on 23 September, putting Gigapixel inside Photoshop and Lightroom on metered credits and leaving legacy perpetual licences unaddressed. It extended its creative tools into Google Gemini and expanded them inside Claude, signed Jet2, was named partner to the NHL's 32 clubs with no launch date given, and now lists Veo 3.1, Kling 3.0, Runway Gen-4.5, Luma Ray3.14 and Seedance 2.0 as partner video models, with the two Chinese models limited to individual plans. ElevenLabs shipped Eleven v4 on 28 September, cloning a voice from 10 seconds of audio across more than 90 languages; it ranks first on the Artificial Analysis arena at 1319 Elo against Cartesia Sonic 3.6 at 1276, and a $300M secondary sale doubled the company's valuation to $22bn. In video, the Sora 2 API was cut off on 24 September with OpenAI naming no replacement. Runway launched an ads agent on 30 September that publishes variants to Meta, Google and TikTok and feeds performance data back into the next round. Pika relaunched as a studio that routes requests across other firms' models. Google took its real-time Gemini avatar agent to general availability on 24 September with SynthID on every stream, and the same day published a four-system orchestration layer chaining Gemini and Veo into sequences of up to 10 minutes. That result is scored on Google's own benchmark, is not packaged as a product and has not been independently replicated. Veo 3.1 clips still cap at 8 seconds.\n\nThe more instructive evidence was negative, and much of it concerned measurement. Incode reported that GenD, a leading public deepfake detector, fell from 91.2% benchmark AUROC to just over 60% on its own identity-verification data. A self-audited detector ensemble returned a 99.85% probability of AI in a case where four of five detectors were silent and the fifth said \"real\". Vidmoat found seven bugs in its own video-editing benchmark, including an agent run marked complete whose export was black for 85 of 92 seconds. Krisp's open benchmark of 265 recordings showed voice isolation cutting pooled word error rate by 73% while making clean phone audio slightly worse, from 3.48% to 3.91%. Atlas Cloud, which sells its own endpoints, found only 3 of 36 cloud image-editing endpoints accept a mask file. On outcomes, AIR Media-Tech, a dubbing vendor with an interest in the answer, reported that YouTube's auto-dubbed tracks held viewers 4 to 10 times less than professional dubs across 400+ channels. A practitioner's paid-media test had 20 fully generated ad variants finish 40% below two hand-finished ads on click-through. VML's survey of 28,000 consumers in 17 countries found 49% say AI product images lower their trust in a brand. A peer-reviewed survey of 105 3D practitioners found generated meshes routinely failing rigging, with over a third of those answering spending 2-5 hours per mesh on topology fixes, even as Tripo P2 shipped native quad output and Meshy researchers reported 94.1% of UV seam predictions unwrapping without post-processing. Roblox opened its prompt-to-game tool as a public alpha in three markets, with roughly 9,000 games published, mostly by first-time Studio users, against a third straight quarter of falling daily users. Nothing in the fortnight changed the overall picture: the new evidence sharpened existing limits and did not move them.\n\n## Key Tensions\n\n- **Generation is cheap; a usable asset is not.** Unit prices have collapsed, but the labour has moved downstream. Practitioner analyses put the keeper rate for short video at roughly one in three, a professional composer found under 1% of more than 20,000 generated tracks production-viable, and game studios report motion-capture cleanup and retargeting consuming 30-50% of animation effort. Review capacity has not scaled with output: one survey found 46% of marketing professionals reporting content stuck in review queues.\n\n- **Supply floods channels that audiences and platforms then close.** Volume and attention have decoupled. Qobuz says AI tracks account for 0.38% of its streams and that it demonetises 60% of AI music, while 221,900 new AI dramas were uploaded in China in the first half of 2026 and 1.3% reached profitability thresholds. Platforms are answering with penalties: Tidal stopped paying royalties on fully generated music from 15 July, YouTube bars template-based AI content from Partner Program monetisation, and TikTok's disclosure toggle carries a 30-40% reach penalty.\n\n- **Platform owners absorb the feature and specialists lose pricing power.** Background music, upscaling, dubbing and style transfer now ship inside Premiere, Photoshop, Gemini and YouTube, often against a shared credit balance. Beatoven.ai's founder cited heavily funded entrants when closing its consumer service in August, and Adobe now owns Topaz. Buyers gain convenience and inherit churn: Runway retired two models on 30 July with no grace period, Photoshop began pulling Gemini 2.5 on 10 September ahead of Google's 2 October deprecation, and aggregator routing can change style, latency or cost without the user touching anything.\n\n- **Provenance is mandated while detection keeps failing.** Marking synthetic media is now a legal duty under EU AI Act Article 50 and California's SB 942, and OpenAI, Anthropic and Google attach C2PA credentials and invisible watermarks by default. Detection does not survive contact with current generators: the DF26 benchmark saw detectors fall from 94% AUC on legacy data to 48-70% on modern full-scene video. Credentials are stripped in distribution and a watermark-removal plugin drew 19 thousand stars in a day, so a missing mark does not show that content is human-made, a limit OpenAI's own documentation states for its tools.\n\n- **Rights and consent are being negotiated after the fact.** Licensing is replacing litigation for those large enough to negotiate: Suno's new models are built with licensed music from Warner, BMG and Believe, though nobody has disclosed what artists will be paid. The output itself remains legally thin, since the US Supreme Court declined to hear Thaler v. Perlmutter on 2 March 2026 and purely generated work stays unregistrable in the US, while Adobe's indemnity excludes claims arising from modifying or combining output. Performers are contesting consent directly: nearly 1,000 actors, agents and others signed an open letter over demands that child actors allow AI use of their voices.\n\n## Top 10 Evidence Items\n\n1. **Adobe Owns Topaz Labs as $340 Million Deal Closes: What Changes for Photographers** (news-coverage) — Shows the platform-owner consolidation pattern the briefing calls the fortnight's clearest trend, with legacy licences left stranded as a cost of that absorption. https://www.photographytalk.com/adobe-owns-topaz-labs/\n2. **Why deepfake detection benchmarks fall apart on production data — Incode** (case-study) — Demonstrates provenance tooling failing exactly where it is mandated to work, undercutting the EU AI Act transparency regime now in force. https://www.incode.com/blog/deepfake-detection-in-production/\n3. **STT handles noise now. It still can’t handle a background voice.** (case-study) — A rare rigorously measured result that complicates the clean capability narrative: the same tool both fixes and degrades audio depending on input. https://krisp.ai/blog/voice-isolation-benchmark/?ref=toolcenter\n4. **ElevenLabs’ new v4 speech model supports more expression control and 90 languages** (news-coverage) — Anchors the money-follows-capability side of the tension, pairing voice-cloning fidelity with the valuation and ARR figures the briefing cites. https://techcrunch.com/2026/09/28/elevenlabs-new-v4-speech-model-supports-more-expression-control-and-90-languages/\n5. **When Should You Stop Automating Creative and Customer Service?** (opinion) — A concrete paid-media test contradicting the cost-savings case for full automation, supporting the keeper-rate and review-capacity tension. https://www.cmswire.com/customer-experience/when-should-you-stop-automating-creative-and-customer-service/\n6. **Is YouTube's AI Dubbing Safe for Your Channel? Auto Dub vs Pro, With Real Retention Data** (case-study) — Direct evidence for the demand-side stall: default-on scaled dubbing actively costs retention, not just quality. https://air.io/en/youtube-hacks/is-youtubes-ai-dubbing-safe-for-your-channel-auto-dub-vs-pro-with-real-retention-data\n7. **Beatoven.ai shuts consumer AI music service as core team joins Rusk Media** (news-coverage) — A named casualty that illustrates platform owners absorbing features and squeezing specialists out of the market. https://routenote.com/radar/beatoven-ai-shuts-consumer-ai-music-service-as-core-team-joins-rusk-media/\n8. **Inside Roblox’s Massive Bet On AI-assisted Creation** (industry-report) — Captures the split between internal AI tooling scaling and external audience acceptance falling, the core tension of the domain. https://naavik.co/ai-gaming/inside-robloxs-massive-bet-on-ai-assisted-creation/\n9. **Topology-based failure modes in AI-generated 3D assets for production rigging pipelines: a practitioner-informed benchmarking framework** (research-paper) — Grounds the 'someone still stands between generation and publication' claim with hard numbers on rigging rework. https://link.springer.com/article/10.1007/s11042-026-21944-w?\n10. **Suno made a new AI music model with major labels. Here's what it means** (news-coverage) — Shows rights being renegotiated after the fact, with licensing substituting for litigation while artist pay stays undisclosed. https://www.latimes.com/entertainment-arts/business/story/2026-09-21/ai-music-licensing-deals-labels-artists-who-benefits",
  "execSummary": "**The headline:** Creative AI is now a feature inside software you already pay for, and the models underneath can vanish. Tests published this fortnight show fully generated ads and dubs still losing to human-finished work.\n\n### The Picture\n\nThe capability question is settled: in blind tests listeners cannot reliably tell generated music from human, and cloned voices fool most people. The acceptance question is not. Most marketing teams now use AI for backgrounds, resizes, variants and training video, with a person approving what ships, and if that describes you, you are in the pack. The better results on record come from teams that train models on their own products and keep human finishing; the exposed ones publish fully generated work to audiences who keep telling surveys they trust it less. Vendors prosper either way: Adobe reports annual recurring revenue from AI-first products above $650 million.\n\n### This Fortnight\n\n- **Adobe closed its roughly $340 million purchase of Topaz Labs on September 23.** Topaz's image-enlarging tool now runs inside Photoshop and Lightroom against a metered credit balance, and Adobe has said nothing about customers who bought Topaz outright under perpetual licenses. Specialist tools are becoming line items in subscriptions you already hold, so ask your creative teams which standalone licenses they depend on and what the credit meter will cost at their volume.\n\n- **OpenAI cut off developer access to its Sora 2 video model on September 24 and named no replacement.** The consumer app had already closed in April after earning $2.1 million in total. Any workflow built on a single video model needs a tested fallback, because a vendor can retire a model without offering anything to move to.\n\n- **ElevenLabs released a model on September 28 that clones a voice from 10 seconds of audio.** A share sale in the same fortnight doubled the company's valuation. Cheaper, faster cloning helps localization and customer service, and it equally helps anyone impersonating your chief financial officer, so payment approvals that rest on recognizing a voice need a second check.\n\n- **Incode reported that a leading public deepfake detector fell from 91.2 percent on academic benchmarks to just over 60 percent on its own identity-verification data.** Incode, which sells identity verification, attributes the gap to the difference between polished public test images and unfiltered real-world selfies. Treat detection software as one fraud signal among several, never as a verdict on whether a face or voice is real.\n\n- **In one paid-media test, 20 fully AI-generated ad variants finished 40 percent below two hand-finished ads on click-through.** The winners paired a model trained on the brand's own products with human copy, retouching and layout, and a dubbing vendor with an interest in the answer reported YouTube's automatic dubs holding viewers far less than professional ones. Budget for human finishing on anything customer-facing, and judge these tools on results, not volume.\n\n### Coming Up\n\n- **Google starts retiring an older Gemini image model on October 2, and Photoshop has already begun removing it.** Adobe's editing features now route work across its own and partner models, so a retirement upstream can change what your team's saved workflows produce. Have someone list which models your recurring creative jobs depend on, and re-test output whenever one changes.\n\n- **EU rules requiring disclosure of AI-generated media have applied since August 2, and the duty sits with the advertiser, not just the agency.** Fines run up to €15 million, and European Commission guidelines treat hidden metadata alone as insufficient. Check before your next European campaign who in your agency contracts is responsible for visible labeling, and watch the first enforcement cases for how strictly regulators read the rule.\n\n- **YouTube plans to pilot real-time dubbing for livestreams in early 2027.** Languages and technology are undisclosed, its existing automatic dubbing is already on by default, and performers' consent terms are still being fought over in courts and union contracts. Before relying on default dubs in a new market, compare viewer retention by language track and keep professional dubbing for humor, children's and emotional content.\n\n### What's Hard About This\n\n- **Generation is cheap, but a publishable asset is not.** One industry assessment found 68 percent of enterprise video deployments still need human intervention to reach publishable quality, and review capacity has not grown with output. The saving is real only if you count the checking, so measure cost per accepted asset, not cost per generation.\n\n- **Your audience judges the output, and much of it is wary.** Gartner finds half of consumers would rather buy from brands that keep AI out of public-facing content, and platforms are cutting reach and payouts for generated material. No tool upgrade fixes that; it is a decision about where in your customer experience synthetic content is acceptable and how it is disclosed.\n\n- **Marking synthetic media is now a legal duty, while detecting it keeps failing against current generators.** Detectors that reach 94 percent on older test footage manage 48 to 70 percent on video from today's tools, and labels are stripped as content moves between platforms. A missing mark does not prove something is human-made, so verifying anything that matters, from a supplier's video call to a news clip, has to rest on process, not software.",
  "headline": "Creative AI is now a feature inside software you already pay for, and the models underneath can vanish. Tests published this fortnight show fully generated ads and dubs still losing to human-finished work.",
  "execSummarySections": [
    {
      "id": "the-picture",
      "title": "The Picture",
      "body": "The capability question is settled: in blind tests listeners cannot reliably tell generated music from human, and cloned voices fool most people. The acceptance question is not. Most marketing teams now use AI for backgrounds, resizes, variants and training video, with a person approving what ships, and if that describes you, you are in the pack. The better results on record come from teams that train models on their own products and keep human finishing; the exposed ones publish fully generated work to audiences who keep telling surveys they trust it less. Vendors prosper either way: Adobe reports annual recurring revenue from AI-first products above $650 million."
    },
    {
      "id": "this-fortnight",
      "title": "This Fortnight",
      "body": "- **Adobe closed its roughly $340 million purchase of Topaz Labs on September 23.** Topaz's image-enlarging tool now runs inside Photoshop and Lightroom against a metered credit balance, and Adobe has said nothing about customers who bought Topaz outright under perpetual licenses. Specialist tools are becoming line items in subscriptions you already hold, so ask your creative teams which standalone licenses they depend on and what the credit meter will cost at their volume.\n\n- **OpenAI cut off developer access to its Sora 2 video model on September 24 and named no replacement.** The consumer app had already closed in April after earning $2.1 million in total. Any workflow built on a single video model needs a tested fallback, because a vendor can retire a model without offering anything to move to.\n\n- **ElevenLabs released a model on September 28 that clones a voice from 10 seconds of audio.** A share sale in the same fortnight doubled the company's valuation. Cheaper, faster cloning helps localization and customer service, and it equally helps anyone impersonating your chief financial officer, so payment approvals that rest on recognizing a voice need a second check.\n\n- **Incode reported that a leading public deepfake detector fell from 91.2 percent on academic benchmarks to just over 60 percent on its own identity-verification data.** Incode, which sells identity verification, attributes the gap to the difference between polished public test images and unfiltered real-world selfies. Treat detection software as one fraud signal among several, never as a verdict on whether a face or voice is real.\n\n- **In one paid-media test, 20 fully AI-generated ad variants finished 40 percent below two hand-finished ads on click-through.** The winners paired a model trained on the brand's own products with human copy, retouching and layout, and a dubbing vendor with an interest in the answer reported YouTube's automatic dubs holding viewers far less than professional ones. Budget for human finishing on anything customer-facing, and judge these tools on results, not volume."
    },
    {
      "id": "coming-up",
      "title": "Coming Up",
      "body": "- **Google starts retiring an older Gemini image model on October 2, and Photoshop has already begun removing it.** Adobe's editing features now route work across its own and partner models, so a retirement upstream can change what your team's saved workflows produce. Have someone list which models your recurring creative jobs depend on, and re-test output whenever one changes.\n\n- **EU rules requiring disclosure of AI-generated media have applied since August 2, and the duty sits with the advertiser, not just the agency.** Fines run up to €15 million, and European Commission guidelines treat hidden metadata alone as insufficient. Check before your next European campaign who in your agency contracts is responsible for visible labeling, and watch the first enforcement cases for how strictly regulators read the rule.\n\n- **YouTube plans to pilot real-time dubbing for livestreams in early 2027.** Languages and technology are undisclosed, its existing automatic dubbing is already on by default, and performers' consent terms are still being fought over in courts and union contracts. Before relying on default dubs in a new market, compare viewer retention by language track and keep professional dubbing for humor, children's and emotional content."
    },
    {
      "id": "whats-hard-about-this",
      "title": "What's Hard About This",
      "body": "- **Generation is cheap, but a publishable asset is not.** One industry assessment found 68 percent of enterprise video deployments still need human intervention to reach publishable quality, and review capacity has not grown with output. The saving is real only if you count the checking, so measure cost per accepted asset, not cost per generation.\n\n- **Your audience judges the output, and much of it is wary.** Gartner finds half of consumers would rather buy from brands that keep AI out of public-facing content, and platforms are cutting reach and payouts for generated material. No tool upgrade fixes that; it is a decision about where in your customer experience synthetic content is acceptable and how it is disclosed.\n\n- **Marking synthetic media is now a legal duty, while detecting it keeps failing against current generators.** Detectors that reach 94 percent on older test footage manage 48 to 70 percent on video from today's tools, and labels are stripped as content moves between platforms. A missing mark does not prove something is human-made, so verifying anything that matters, from a supplier's video call to a news clip, has to rest on process, not software."
    }
  ],
  "url": "https://www.thestateofplay.ai/domain/creative-generative-media",
  "license": "CC BY 4.0",
  "licenseUrl": "https://creativecommons.org/licenses/by/4.0/",
  "generatedAt": "2026-10-01"
}