Nerra Network

Archives
Log in
Subscribe
August 2, 2026

You can now feed an AI dozens of reference clips and… · M&A Beginners 🎓

View this email in your browser
Models & Agents for Beginners — AI explained simply — for beginners and teens.

Models & Agents for Beginners

AI explained simply — for beginners and teens.

Ep 121 · Aug 2, 2026

🎧 Today's episode
Episode 121 · You can now feed an AI dozens of reference clips and get a full 30-second video with sound in one shot.
2026-08-02
▶ Listen now
You can now feed an AI dozens of reference clips and get a full 30-second video with sound in one shot. ByteDance just released Seedance 2.5, an AI video model that produces up to 30-second clips complete with built-in audio from dozens of reference images, videos, and audio files. The system handles both visuals and sound together instead of requiring separate editing steps, which is three times the length of what Google’s Gemini Omni Flash currently generates. This matters because it removes the need to cut together multiple short clips one at a time, opening the door for faster creative work on school projects, social media, or hobby videos. Today we’ll look at exactly how Seedance 2.5 works, why older prompting tricks are fading, and two hands-on experiments you can run right now without any special software.

The Big Story

ByteDance released Seedance 2.5, an AI system that accepts dozens of reference images, videos, and audio files and produces a single finished 30-second video clip that already contains matching sound.

The model generates both the moving images and the audio track in one pass rather than forcing users to add sound in a second program. Think of it like giving a friend a pile of vacation photos, short phone clips, and voice notes and asking them to create one complete short film while also composing and recording the background music at the same time. Because the system works with up to 30 seconds of output, it is three times longer than the clips currently produced by Google’s Gemini Omni Flash.

Users can supply dozens of separate reference files, letting the model draw visual style, motion, and audio cues from many different sources at once. For advertising teams this removes the old workflow of stitching together many short clips manually. A student making a field-trip summary could upload photos from the day, add a voice note, and receive a polished 30-second explanation video without learning video-editing software. A hobby creator testing TikTok ideas could generate several versions in the time it used to take to finish one.

The change lowers the barrier between having an idea and sharing a finished piece of content. If you enjoy making short videos for class presentations or personal projects, tools like this let you experiment more often without spending hours on post-production. The article notes that ad teams in particular could see their production process speed up dramatically because one prompt now replaces multiple editing steps.

You can test similar reference-based video generation today inside the free Gemini app by uploading two or three photos and typing a short prompt such as “make a 10-second clip of this scene with calm background music.” Start with simple references and watch how the model combines your inputs into one short scene with sound. Source: the-decoder.com


Explain Like I'm 14

You know how in math class the teacher used to make you write out every single step of a long problem even though you already knew the answer?

Back in summer 2024, people used the same trick with AI by typing the words “think step by step” so the model would break a question into smaller pieces and check its own work along the way. That extra sentence often improved the final answer because the model was still learning how to keep track of multiple steps at once.

Newer models have now been trained on so many examples that they already perform those intermediate checks inside their own processing. When researchers tested the same “think step by step” instruction on today’s models, it no longer produced better results because the models were already doing the step-by-step work automatically.

The shift happened because the training data grew large enough for the models to absorb multi-step reasoning as a built-in habit rather than something that needed an external reminder. What used to be an extra prompt has become part of how the model thinks by default.

So when you hear that chain-of-thought prompting no longer helps, it simply means the model has practiced enough problems that the old reminder is no longer necessary. Source: x.com


Cool Stuff & Try This

Make an old-school meme that looks like it came from 2012 The AI Meme Generator prompt shared on Reddit’s r/ChatGPT community lets you create a funny, slightly blurry meme that feels like something you would have scrolled past on Facebook years ago. It uses a regular candid photo style with JPEG compression and bold text instead of clean modern graphics. This is useful if you want to understand how detailed prompts control image quality and tone, or if you just like sharing quick jokes with friends. Go to reddit.com/r/ChatGPT, search for the post titled “AI Meme Generator,” copy the full prompt, and paste it into any free chatbot such as ChatGPT or Claude. Then replace the example topic with something specific like “when the cafeteria runs out of pizza on pizza day” and generate the image to see how the AI keeps the awkward, low-resolution look. Source: reddit.com

One place to chat with many different AIs at once A new dashboard service gives access to more than twenty AI models including ChatGPT, Claude, and Gemini inside a single interface for a flat monthly fee of $79. You can switch between models instantly and compare how each one answers the same question without opening multiple tabs or apps. This is helpful when you want to see different writing styles or pick the answer that feels clearest for a homework question. The Mashable article linked below lists current signup details and pricing. Once you have access, type the same simple question into two different models side by side and notice how their sentence structure and level of detail change. Source: Google News


Quick Bits

AI music generator ruled to have copied songs A Munich court ruled that the AI music tool Suno violated copyright by both storing six specific songs inside its model during training and reproducing parts of them in new outputs. The court rejected Germany’s text-and-data-mining exception as well as the U.S. fair-use defense, showing that legal questions about AI training data are still being decided case by case. Source: the-decoder.com

Students asking chatbots for college advice High school students are increasingly turning to AI chatbots for help choosing classes, writing applications, and planning college visits instead of speaking with human counselors. The trend raises practical questions about how much personal guidance should come from a program versus a trained advisor who knows the student’s full situation. Source: Google News

💬 Reply to this email — Patrick reads every one.

Share: X · LinkedIn · WhatsApp

Forwarded this email? Subscribe here — it's free.

▶ Listen to the podcast

📺 Watch on YouTube  ·  📝 Read the blog  ·  🖼 Free image gallery (CC BY-SA)  ·  📊 Data Hub & Story Trackers  ·  🧭 Start Here

Nerra Network · AI-narrated voice (Grok TTS) · Editorial by Patrick

You're receiving this because you subscribed to Models & Agents for Beginners on nerranetwork.com.

Issue #121 · Models & Agents for Beginners · Aug 2, 2026
Don't miss what's next. Subscribe to Nerra Network:
← Newer Mortgage rates at one-year highs mean Canadian… · MIT 📈 Older → OpenAI’s internal Astra model solved ten decade-old… · M&A 🤖
nerranetwork.com
Powered by Buttondown, the easiest way to start and grow your newsletter.