Thanks to Jackson Beaman & crew for putting together a great event yesterday in SF. I joined him, KD Deshpande (founder of Simplified), and Sofiia Shvets (founder of Let’s Enhance & Claid.ai) for a 20-minute panel discussion (which starts at 3:32:03 or so, in case the embedded version doesn’t jump you to the proper spot) about creating production-ready imagery using AI. Enjoy, and please let me know if you have any comments or questions!
A little weekend drone-NeRF fun
Thanks as always to the guys at Luma Labs for making it so ridiculously easy to generate 3D scenes from simple orbits:
The Founding Fathers talk AI art
Well, not exactly—but T-Paine’s words about how we value things still resonate today:

We humans are fairly good at pricing effort (notably in dollars paid per hour worked), but we struggle much more with pricing value. Cue the possibly apocryphal story about Picasso asking $10,000 for a drawing he sketched in a matter of seconds, but the ability to create which had taken him a lifetime.
A couple of related thoughts:
- My artist friend is a former Olympic athlete who talks about how people bond through shared struggle, particularly in athletics. For him, someone using AI-powered tools is similar to a guy showing up at the gym with a forklift, using it to move a bunch of weight, and then wanting to bond afterwards with the actual weightlifters.
- I see ostensible thought leaders crowing about the importance of “taste,” but I wonder how they think that taste is or will be developed in the absence of effort.
- As was said of—and by?—Steve Jobs, “The journey is the reward.”
[Via Louis DeScioli]
After Effects + Midjourney + Runway = Harry Potter magic
It’s bonkers what one person can now create—bonkers!
I edited out ziplines to make a Harry Potter flying video, added something special at the end
byu/moviemaker887 inAfterEffects
I took a video of a guy zip lining in full Harry Potter costume and edited out the zip lines to make it look like he was flying. I mainly used Content Aware Fill and the free Redgiant/Maxon script 3D Plane Stamp to achieve this.
For the surprise bit at the end, I used Midjourney and Runway’s Motion Brush to generate and animate the clothing.
Trapcode Particular was used for the rain in the final shot.
I also did a full sky replacement in each shot and used assets from ProductionCrate for the lighting and magic wand blast.
[Via Victoria Nece]
Krea upgrades its realtime generation
I had the pleasure of hanging out with these crazy-fast-moving guys last week, and I remain amazed at the speed of their shipping velocity. Check out the latest updates to their realtime canvas:
big upgrades to quality!
announcing Portrait, Concept, CGI, and Cartoon.
try them for free in KREA real-time (link below). pic.twitter.com/iLqyWHT1Vn
— KREA AI (@krea_ai) January 25, 2024
Check out how trailblazing artist Martin Nebelong is putting it to use:
Speaking about control.. another step in the right direction!
Testing out the latest Krea ai updates.. crazy stuff
Combined with some Magnific ai magic.#ai #art pic.twitter.com/9Zz4JY4W15
— Martin Nebelong (@MartinNebelong) January 26, 2024
“Fonts hanging out”
Hah—serifs are “little top hats & booties.” Enjoy a bit of inspired typographical silliness:
Google introduces Lumiere for video generation & editing
Man, not a day goes by without the arrival of some new & mind-blowing magic—not a day!
We introduce Lumiere — a text-to-video diffusion model designed for synthesizing videos that portray realistic, diverse and coherent motion — a pivotal challenge in video synthesis. To this end, we introduce a Space-Time U-Net architecture that generates the entire temporal duration of the video at once, through a single pass in the model. This is in contrast to existing video models which synthesize distant keyframes followed by temporal super-resolution — an approach that inherently makes global temporal consistency difficult to achieve. […]
We demonstrate state-of-the-art text-to-video generation results, and show that our design easily facilitates a wide range of content creation tasks and video editing applications, including image-to-video, video inpainting, and stylized generation.
Content credentials are coming to DALL•E
From its first launch, Adobe Firefly has included support for content credentials, providing more transparency around the origin of generated images, and I’m very pleased to see Open AI moving in the same direction:
Early this year, we will implement the Coalition for Content Provenance and Authenticity’s digital credentials—an approach that encodes details about the content’s provenance using cryptography—for images generated by DALL·E 3.
We are also experimenting with a provenance classifier, a new tool for detecting images generated by DALL·E. Our internal testing has shown promising early results, even where images have been subject to common types of modifications. We plan to soon make it available to our first group of testers—including journalists, platforms, and researchers—for feedback.
Adobe Announces Inaugural Film & TV Fund, Committing $6 Million to Support Underrepresented Creators
In her 12+ years in Adobe’s video group, my wife Margot worked to bring more women into the world of editing & filmmaking, participating in efforts supporting all kinds of filmmakers across a diverse range of ages, genders, types of subject matter, experience levels, and backgrounds. I’m delighted to see such efforts continuing & growing:
Adobe and the Adobe Foundation will partner with a cohort of global organizations that are committed to empowering underrepresented communities, including Easterseals, Gold House, Latinx House, Sundance Institute and Yuvaa, funding fellowships and apprenticeships that offer direct, hands-on industry access. The grants will also enable organizations to directly support filmmakers in their communities with funding for short and feature films.
The first fellowship is a collaboration with the NAACP, designed to increase representation in post-production. The NAACP Editing Fellowship is a 14-week program focused on education and training, career growth and workplace experience and will include access to Adobe Creative Cloud to further set up emerging creators with the necessary tools. Applications open on Jan. 18, with four fellows selected to participate in the program starting in May.
Premiere Pro ups its audio game
“If you want to make a movie look good, make it sound good.” That’s the spirit in which Adobe is introducing a wide range of enhancements to audio handling in Premiere Pro:
According to the team, the audio workflow changes now available in the beta include:
- Interactive Fade Handles: Now you can simply click and drag from the edge of a clip to create a variety of custom audio fades in the timeline or drag across two clips to create a crossfade. These visual fades provide more precision and control over audio transitions while making it easy to see where they are applied across your sequence.
- AI-powered Audio Category Tagging: When you drag clips into the sequence, they’ll automatically be identified and labeled with new icons for dialogue, music, sound effects, or ambience. A single click on the icon provides access to the most relevant tools for that audio type in the Essential Sound panel — such as Loudness Matching or Auto Ducking.
- Redesigned FX Clip Badges: An updated badge makes it easier for you to see which clips have effects added to them. New effects can be added by right clicking the badge, and a single click opens the Effect Control panel for even more adjustment without changing the workspace or searching for the panel.
- Modern, Intelligent Waveforms and Clips: Waveforms now dynamically resize when you change the track height and improved clip colors make it easier for you to see and work with audio on the timeline.
Tutorial: Firefly + Character Animator
Helping discover Dave Werner & bring him into Adobe remains one of my favorite accomplishments at the company. He continues to do great work in designing characters as well as the tools that can bring them to life. Watch how he combines Firefly with Adobe Character Animator to create & animate a stylish tiger:
Adobe Firefly’s text to image feature lets you generate imaginative characters and assets with AI. But what if you want to turn them into animated characters with performance capture and control over elements like arm movements, pupils, talking, and more? In this tutorial, we’ll walk through the process of taking a static Adobe Firefly character and turning it into an animated puppet using Adobe Photoshop or Illustrator plus Character Animator.
Six Ways to Spice Up Your Photos in Lightroom
Adding glows & fog, making bokeh balls, and more—lots of nice, bite-sized demos, conveniently navigable by chapter (hover over the progress bar):
“How Adobe is managing the AI copyright dilemma, with general counsel Dana Rao”
Honestly, if you asked, “Hey, wanna spend an hour+ listening to current and former intellectual property attorneys talking about EU antitrust regulation, ethical data sourcing, and digital provenance,” I might say, “Ehmm, I’m good!”—but Nilay Patel & Dana Rao make it work.
I found the conversation surprisingly engrossing & fast-moving, and I was really happy to hear Dana (with whom I’ve gotten to work some regarding AI ethics) share thoughtful insights into how the company forms its perspectives & works to put its values into practice. I think you’ll enjoy it—perhaps more than you’d expect!
FPV Miniatur Wunderland
Insta360 takes us down & not so dirty around Hamburg’s Miniatur Wunderland in this fun 2-minute tour:
Deeply chill photography
(Cue Metallica’s Trapped Under Ice!)
Russell Brown & some of my old Photoshop teammates recently ventured into -40º (!!) weather in Canada, pushing themselves & their gear to the limits to witness & capture the Northern Lights:
Perhaps on future trips they can team up with these folks:
To film an ice hockey match from this new angle of action, Axis Communications used a discrete modular camera — commonly seen in ATM machines, onboard vehicles, and other small spaces where a tiny camera needs to fit — and froze it inside the ice.
Check out the results:
Behind—and under—the scenes:
Adobe’s hiring a prototyper to explore generative AI
We’re only just beginning to discover the experiential possibilities around generative creation, so I’m excited to see this rare gig open up:
You will build new and innovative user interactions and interfaces geared towards our customers unique needs, test and refine those interfaces in collaboration with academic research, user researchers, designers, artists and product teams.
Check out the listing for the full details.
Amazing Lego recreations of extreme sports
Stunning work from Legosteeze. Make sure to click the arrows on the post to see all the clips (amazingly based on real-world footage!):
View this post on Instagram
View this post on Instagram
[Via Cristobal Garcia]
Adobe Project Primrose dazzles Sofía Vergara
My friend Kevin had the honor of designing the designing the art & animation for this interactive wearable demo. Stick around (or jump) to the end to see the moving images:
@el_hormiguero Marron nos trae el proyecto Primrose desarrollado por @Adobe: el vestido interactivo capaz de cambiar su apariencia #vestidointeractivo #adobe #elhormiguero #SofíaVergaraEH ♬ sonido original – El Hormiguero
Two quotes worth reflecting on as we go into the new year
One, I swear I think of this observation from author Sebastian Junger at least once a day:

We’d do well to reflect on it in how we treat our colleagues, and especially—in this time of disruptive AI—how we treat the sensitive, hardworking creators who’ve traditionally supported toolmarkers like Adobe. Our “empowering” tech can all too easily make people feel devalued, thrown away like an old piece of fruit. And when that happens, we’re next.
Two, this observation hits me where I live:

I’ve joked for years about my “Irish Alzheimer’s,” in which one forgets everything but the grudges. It’s funny ’cause it’s true—but taken any real distance (focusing on failures & futility), it becomes corrosive, “like taking poison and hoping the other guy gets sick.”
Earlier today an old friend observed, “I’ve always had a justice hang-up.” So have I, and that’s part of what made us friends for so long.
But as I told him, “It’s such a double-edged sword: my over-inflamed sense of justice is a lot of what causes me to speak up too sharply and then light my way by all the burning bridges.” Finding the balance—between apathetic acquiescence on one end & alienating militancy on the other—can be hard.
So, for 2024 I’m trying to lead with gratitude. It’s the best antidote, I’m finding, to bitterness & bile. Let’s be glad for our fleeting opportunities to do, as Mother Teresa put it, “small things with great love.”
Here’s to courage, empathy, and wisdom for our year ahead.
Adobe Firefly named “Product of the Year”
Nice props from The Futurum Group:
Here is why: Adobe Firefly is the most commercially successful generative AI product ever launched. Since it was introduced in March in beta and made generally available in June, at last count in October, Firefly users have generated more than 3 billion images. Adobe says Firefly has attracted a significant number of new Adobe users, making it hard to imagine that Firefly is not aiding Adobe’s bottom line.
Happy New Year!
Hey gang—here’s to having a great 2024 of making the world more beautiful & fun. Here’s a little 3D creation (with processing courtesy of Luma Labs) made from some New Year’s Eve drone footage I captured at Gaviota State Beach. (If it’s not loading for some reason, you can see a video version in this tweet).
AI Holiday Leftovers, Vol. 3
- Fun with famous IP:
- Vectors: StarVector: Generating Scalable Vector Graphics Code from Images
- How to train a custom Stable Diffusion model to generate consistent characters via LensGo.ai, which lets you train 3 custom models for free every month.
- Insane 128x zoom-in on AI-generated meat.
AI Holiday Leftovers, Vol. 2
- 3D:
- Paint Anything 3D with Lighting-Less Texture Diffusion Models: “Paint3D is a novel coarse-to-fine generative framework that is capable of producing high-resolution, lighting-less, and diverse 2K UV texture maps for untextured 3D meshes conditioned on text or image inputs.”
- “Google just revealed an ABSOLUTE depth estimation model. As opposed to recent depth models (Marigold, PatchFusion) which aim for maximum details, DMD aims to estimate the ABSOLUTE depth (in meters) within the image.”
- Typography:
- Retro-futuristic alphabet rendered with Midjourney V6: “Just swapped out the letter and kept everything else the same. Prompt: Letter “A”, cyberpunk style, metal, retro-futuristic, star wars, intrinsic details, plain black background. Just change the letter only. Not all renders are perfect, some I had to do a few times to get a good match. Try this strategy for any type of cool alphabet!”
- As many others have noted, Midjourney is now good at type. Find more here.
AI Holiday Leftovers, Vol. 1
Dig in, friends. 🙂
- Drawing/painting:
- Using a simple kids’ drawing tablet to create art: “I used @Vizcom_ai to transform the initial sketch. This tool has gotten soo good by now. I then used @LeonardoAi_’s image to image to enhance the initial image a bit, and then used their new motion feature to make it move. I also used @Magnific_AI to add additional details to a few of the images and Decohere AI’s video feature.”
- Latte art: “Photoshop paint sent to @freepik’s live canvas. The first few seconds of the video are real-time to show you how responsive it is. The music was made with @suno_ai_. Animation with Runways Gen-2.”
- Photo editing:
- Google Photos gets a generative upgrade: “Magic Eraser now uses gen AI to fill in detail when users remove unwanted objects from photos. Google Research worked on the MaskGIT generative image transformer for inpainting, and improved segmentation to include shadows and objects attached to people.”
- Clothing/try-on:
- PICTURE: PhotorealistIC virtual Try-on from UnconstRained dEsigns: “We propose a novel virtual try-on from unconstrained designs (ucVTON) task to enable photorealistic synthesis of personalized composite clothing on input human images.”
- AnyDoor is “a diffusion-based image generator with the power to teleport target objects to new scenes at user-specified locations in a harmonious way.”
- SDXL Auto FaceSwap enables to create new images using the face of a source image (example attached).

AI: Tons of recent rad things
- Realtime:
- Oh look, I’m George Clooney! Kinda. You can be, too. FAL AI promises “AI inference faster than you can type.”
- “100ms image generation at 1024×1024. Announcing Segmind-Vega and Segmind-VegaRT, the fastest and smallest, open source models for image generation at the highest resolution.”
- Krea has announced their open beta, “free for everyone.”
- How incredible would it be to have realtime generative brushes like this?
- Drawing to Video, made using Vizcom -> Leonardo -> Pika.
- 3D generation:
- ByteDance has released ImageDream (image to 3D)
- SceneWiz3D offers “A new approach to create high-fidelity 3D scenes from text and 3D object control”
- Image -> depth -> geometry using Marigold + Blender
- 3D for fashion, sculpting, and more:
- This is what Adobe Substance & a notional 3D mode of Firefly Text-to-Image should feel like.
- Outfit Anyone + Animate Anyone = virtual try on + movement.
- Sculpting/rendering via Adobe Substance 3D Modeler + Dreams + Unbound + Krea.
- AnimateDiff v3 was just released.
- Instagram has enabled image generation inside chat (pretty “meh,” in my experience so far), and in stories creation, “It allows you to replace a background of an image into whatever AI generated image you’d like.”
- “Did you know that you can train an AI Art model and get paid every time someone uses it? That’s Generaitiv’s Model Royalties System for you.”
How-to: Combining Photoshop + ComfyUI
It’s a little nerdy even for my blood, but some of my teammates swear by these techniques that enable connecting Photoshop to a hosted instance of Stable Diffusion, enabling one to guide the process via a Photoshop doc and/or custom-trained styles:
“I Draw Better Than AI!”
Hah—I can dig this finger-rich pin from Pictoplasma.

To the moon! Insta360 makes a satellite
What if your tiny planet—a visual genre I’ve enjoyed beating halfway into the ground—were our actual planet? Insta360, on whom I’ve spent crazy amounts of money buying brilliant-if-maddening gear, has now sent their devices to the edge of space:

Catch up on great new Illustrator features in 60 seconds
Let’s talk vector generation, 3D mockup support, image-to-type, and more. Take ‘er away, Deke!
Promising 3D research from Adobe
AI image generation is getting *crazy* fast
Gemini is bonkers
I mean, seriously, what even is all this?? I can’t explain; just please watch.
- 0:00 Intro
- 0:19 Multimodal Dialogue
- 1:32 Multilinguality 2:04
- Game Creation 2:31
- Visual Puzzles 3:17
- Making Connections
- 3:39 Image & Text Generation
- 4:06 Logic & Spatial Reasoning
- 4:55 Translating Visuals
- 5:27 Cultural Understanding
Baby, You Can Drive My Bricks
I’ve had way too much fun creating custom Lego sets based on friends’ & family’s rides, so to help others do it, I’ve made my first custom GPT, “Baby You Can Drive My Bricks.” Take it for a spin & let me know what you create!

Pika Labs “Idea-to-Video” looks stunning
It’s ludicrous to think that these folks formed the company just six months ago, and even more ludicrous to see what the model can already do—from video synthesis, to image animation, to inpainting/outpainting:
Our vision for Pika is to enable everyone to be the director of their own stories and to bring out the creator in each of us. Today, we reached a milestone that brings us closer to our vision. We are thrilled to unveil Pika 1.0, a major product upgrade that includes a new AI model capable of generating and editing videos in diverse styles such as 3D animation, anime, cartoon and cinematic, and a new web experience that makes it easier to use. You can join the waitlist for Pika 1.0 at https://pika.art.
“Emu Edit” enables instructional image editing
This tech—or something much like it—is going to be a very BFD. Imagine simply describing the change you’d like to see in your image—and then seeing it.
[Generative models] still face limitations when it comes to offering precise control. That’s why we’re introducing Emu Edit, a novel approach that aims to streamline various image manipulation tasks and bring enhanced capabilities and precision to image editing.
Emu Edit is capable of free-form editing through instructions, encompassing tasks such as local and global editing, removing and adding a background, color and geometry transformations, detection and segmentation, and more. […]
Emu Edit precisely follows instructions, ensuring that pixels in the input image unrelated to the instructions remain untouched. For instance, when adding the text “Aloha!” to a baseball cap, the cap itself should remain unchanged.
And for some conceptually related (but technically distinct) ideas, see previous: Iterative creation with ChatGPT.
“We’re on a mission from God…”
On the off chance you missed me over the last week or so, it’s due to my being off in Illinois with the fam, having fun making silliness like this:
NBA goes NeRF
Here’s a great look at how the scrappy team behind Luma.ai has helped enable beautiful volumetric captures of Phoenix Suns players soaring through the air:
Go behind the scenes of the innovative collaboration between Profectum Media and the Phoenix Suns to discover how we overcame technological and creative challenges to produce the first 3D bullet time neural radiance field NeRF effect in a major sports NBA arena video. This involved not just custom-building a 48 GoPro multi-cam volumetric rig but also integrating advanced AI tools from Luma AI to capture athletes in stunning, frozen-in-time 3D visual sequences. This venture is more than just a glimpse behind the scenes – it’s a peek into the evolving world of sports entertainment and the future of spatial capture.
Phat Splats
If you keep hearing about “Gaussian Splatting” & wondering “WTAF,” check out this nice primer from my buddy Bilawal:
There’s also Two-Minute Papers, offering a characteristically charming & accessible overview:
GenAI demos from Russell Brown
It’s always great to learn from the master—especially when he’s making “spaghetti western” literal!
- The power of selections with Generative Fill
- Create watercolors and other art styles with Generative Fill
- Manage the stacking order of Generative layers

Iterative creation with ChatGPT
I’m really digging the experience of (optionally) taking a photo, feeding it into ChatGPT, and then riffing my way towards an interesting visual outcome. Here’s a gallery in which you can see some of the journeys I’ve undertaken recently.
- Image->description->image quality is often pretty hit-or-miss. Even so, it’s such a compelling possibility that I keep wanting to try it (e.g. seeing a leaf on the ground, wanting to try turning it into a stingray).
- The system attempts to maintain various image properties (e.g. pose, color, style) while varying others (e.g. turning the attached vehicle from a box truck to a tanker while maintaining its general orientation plus specifics like featuring three Holstein cows).
- Overall text creation is vastly improved vs. previous models, though it can still derail. It’s striking that one can iteratively improve a particular line of text (e.g. “Make sure that the second line says ‘TRAIN’“).


GenFill vs. eternal dog-pant mysteries
Hah! This is my kind of ridiculous Adobe social content. 🙂 Happy Friday.
The Young & The Spiderverse
Man, I’m inspired—and TBH a little jealous—seeing 14yo creator Preston Mutanga creating amazing 3D animations, as he’s apparently been doing for fully half his life. I think you’ll enjoy the short talk he gave covering his passions:
The presentation will take the audience on a journey, a journey across the Spider-Verse where a self-taught, young, talented 14-year-old kid used Blender, to create high-quality LEGO animations of movie trailers. Through the use of social media, this young artist’s passion and skill caught the attention of Hollywood producers, leading to a life-changing invitation to animate in a new Hollywood movie.
Hands up for Res Up ⬆️
Speaking of increasing resolution, check out this sneak peek from Adobe MAX:
It’s a video upscaling tool that uses diffusion-based technology and artificial intelligence to convert low-resolution videos to high-resolution videos for applications. Users can directly upscale low-resolution videos to high resolution. They can also zoom-in and crop videos and upscale them to full resolution with high-fidelity visual details and temporal consistency. This is great for those looking to bring new life into older videos or to prevent blurry videos when playing scaled versions on HD screens.
Adventures in Upsampling
Interesting recent finds:
- Google Zoom Enhance. “Using generative AI, Zoom Enhance intelligently fills in the gaps between pixels and predicts fine details, opening up more possibilities when it comes to framing and flexibility to focus on the most important part of your photo.”
- Nick St. Pierre writes, “I just upscaled an image in MJ by 4x, then used Topaz Photo AI to upscale that by another 6x. The final image is 682MP and 32000×21333 pixels large.”
- Here’s a thread of 10 Midjourney upsampling examples, including a direct comparison against Topaz.
Demos: Photoshop Generative AI tips
Demos: Using Generative AI in Illustrator
If you’ve been sleeping on Text to Vector, check out this handful of quick how-to vids that’ll get you up to speed:
- Welcome to Generative AI in Illustrator
- Generate artwork from text with Text to Vector Graphic (Beta)
- Explore creating stunning patterns with Text to Vector Graphics
- Tips for making your best artwork with Text to Vector Graphic (Beta)
- Tips: Take Your Text to Vector Graphic (Beta) patterns to “Wow!”
- Tip: Control your pattern color with Text to Vector Graphic (Beta)
Ai + AI FTW
Check out this quick demo of Illustrator’s new text-to-vector & mockup tools working together:
AI generated Logos onto any surface. pic.twitter.com/qY4tEkVK0Q
— Riley Brown (@rileybrown_ai) October 29, 2023
360º AI: Skybox adds new sketching & style features
Directly sketch inside a 360º canvas, then generate results:
And see also the styles these folks are working to bring online:
Sneak peek: Project Glyph Ease
Easy as ABC, 123?
Project Glyph Ease uses generative AI to create stylized and customized letters in vector format, which can later be used and edited. All a designer needs to do is create three reference letters in a chosen style from existing vector shapes or ones they hand draw on paper, and this technology automatically create the remaining letters in a consistent style. Once created, designers have flexibility to edit the new font since the letters will appear as live text that can be scaled, rotated or moved in the project.
DreamCraft 2D->3D tech looks wild
Can you imagine something like this running in Photoshop, making it possible to re-pose objects and then merge them back into one’s scene?