如何制作AI动画
How an AI Anime Is Created

原始链接: https://www.aventos.dev/our-process?view=blog

该工作室旨在通过利用人工智能降低成本、绕过日本剥削性的制作委员会制度,并为艺术家提供更优厚的薪酬,从而使动画制作在美国实现商业可行性。他们认为,通过简化制作流程,可以实现动画创作的民主化,并培育本土产业。 他们的工作流优先考虑故事的“核心”——即关键的情感或叙事钩子,通常会删减次要情节以确保节奏紧凑且有效。在剧本创作上,他们摒弃了通用的语言大模型,而是采用以人为本的方法进行分镜和节拍表设计。 制作流程遵循精简模式: 1. **初稿:** 生成粗略片段以测试节奏、氛围和角色动作。 2. **AI 动画:** 初稿确定后,利用一致的角色/环境参考以高质量重新生成镜头。 3. **后期制作:** 通过 Photoshop 等工具手动清理 AI 伪影,并对视觉效果进行升频。 4. **音频:** 对话由人类或 AI 完成,音乐和音效则由 AI 生成以提高效率。 归根结底,该工作室将 AI 视为连接创意抱负与预算限制之间的桥梁,在不牺牲核心叙事体验的前提下,实现更快速的迭代。

Hacker News 上关于一篇 AI 生成动画文章(《AI 动画是如何制作的》)的讨论,突显了该媒介在技术进步与伦理争议并存的现状。 评论者认为,尽管 AI 视频生成技术正在不断改进,甚至已能有效表现快节奏的动作场面,但仍存在明显的局限性,例如剪辑破碎以及视觉瑕疵(如角色与物体融合)。虽然有人认为该技术目前适用于商业广告或电子游戏过场动画等短片内容,但也有人反驳称,制作高质量的长篇动画依然遥不可及。 讨论帖还涉及了 AI 对劳动力市场的影响。用户们争论 AI 是否真的能提高人类画师的待遇,许多人指出,动画行业低薪的根源在于创意领域常见的“激情剥削”模式。一小部分用户出于道德立场彻底否定该技术,将整个生成过程斥为剽窃。
相关文章

原文

The story is the most important part of an anime. Bad story-bad show that's it.

Why We Do This

Why not just hire real animators? Why skimp out on anything at all?

The core reason: Entertainment is constrained financially, and that shapes creative decisions. When Hollywood underperforms, we see more layoffs and in our view more generic shows. Our bet is that if we can make anime financially feasible, we can do two things.

One, fans here actually get to make anime in the US. Instead of needing to fly to Japan, speak Japanese, and preferably be Japanese to break into the industry.

Two, pay the employees(artists) better. Look up the production committee system in Japan and how it treats artists if you want the details.

But TLDR; anime studios are broke, the system isn't changing anytime soon, and it pushes independent contractors into working for dirt poor wages.

We think a an AI studio fixes many problems in the industry. Smaller and faster means more financially feasible so we can give better wages for artists and enable a new industry to thrive in the US.

1. Story Writing & Adaptation

When we choose a story, or write one ourselves, we always have to think about the consumer. That usually means writing for the standard anime watcher, typically Gen Z or Gen Alpha, around universal experiences like high school and exams. Most stories we write target a younger audience, male leaning(we're dudes) and follows similar problems and struggles we had when we were younger.

If there's an existing story, like a Korean comic, we have to figure out how to adapt it into an anime. For us, it's quite a big of pre production work because there's just so much content and detail in a story, and we have to bring it down to a 20 minute episode. Otherwise the pacing gets too slow and we never get to the juicy scenes (the fight scenes, the first kiss,etc).

Before anything gets prompted or we start thinking about shots, someone reads the whole source straight through, or at least what counts as one season. We're hunting for the one thing that made a reader keep scrolling at 2am. Why did they like this story? Good visuals? Good writing? Good character relationships?

Whatever it is, that's the "meat" the adaptation gets built around. Everything else, subplots, side characters, slower arcs, gets tested against it: does this scene serve the meat, or is it just there because the original had room for it? If the latter, we cut that content. Less is more.

We often have to change pacing and density too. When you read or scroll through a comic, you go at whatever pace suits you. We force everyone to go at a pace we decide due to the nature of the animation medium, and some people are inevitably going to feel like a story is too slow or too fast either way. So we try to protect the core story-that might mean cutting side characters or skipping subplots entirely. In adapation work, it's an unavoidable part-budgets do not get to explore every single side story.

How we use LLMs here

We don't. It's really, really bad. Super mediocre, super generic. It's not even slop, it's more like bad Hollywood writing combined with Chatgpt 3.

Beat sheet reference frame for an adapted episode
A screenshot of a beat sheet for our animated music video. You can't really beat google docs

2. The First Draft

Once we're happy with the story, we build something we call a "beat sheet." It's another script, but this one details the action and the shot, instead of a broader story. Note that calling it a beat sheet is technically wrong, but it stuck because we're naming the 'beats' of a story.

After the beat sheet, we get a first draft going. We figure out what the characters look like, then use a cheap model to generate fast, rough clips just to see the pacing of the animation.

Making the animation itself as a first draft lets us iterate quickly on the action and the feel of each scene. Does it fit? Does the emotion land? Does it make sense for the character to do this? And it allows us to get the visuals and styling of a episode. What does it look like. What mood do we want to evoke with the lighting, etc.

Screen shot of video editor having the first draft for a story
First draft is simple, we put clips, cut clips, and extend clips to match the vibe of a scene

3. Animation

AI Animation

Once we're happy with the first draft, we take every shot and regenerate it in 720 quality. For the animation, we use a variety of prompting techniques including character references, environment references, extending videos, etc to get the same character and environmental consistency.

The hard part, getting the draft right, is already done, so this ends up being the least time consuming stage. We literally have an agent rerun all the earlier prompts for us, and we only check in to rerun the occasional shot that didn't follow the draft.

4. Post Production

Raw generations get upscaled and cleaned up

Upscaling means making the resolution bettter. While we generate at 720p, most platforms expect 1080p, so we bump up the resolution using DaVinci Resolve or Topaz Labs (now part of Adobe).

Cleaning up is a newer task that AI animation studios specifically have to deal with. Pause on almost any AI generated clip and you'll spot something off: colors morphing strangely, lighting that doesn't stay consistent, mouths moving in ways that don't quite make sense. Fixing that is our job, mostly done by opening photoshop/photo editor and fixing these AI artifacts.

There's also a handful of things we still have to handle(detailed below in post like sfx and voices)

Screen shot of video editor having the final draft for a story
The final draft looks scary on the video editor, but it's just audio tracks and and some extra video tracks. Trust me it's simpler than it looks
  1. Voices

    Dialogue gets recorded/prompted and tested against visuals that are now locked in. A human voice actor records it in a session, or an AI voice generates the performance, depending on the character, the budget, and the timeline. Either way, it has to feel emotionally right. Two laughs from the same character shouldn't sound the same twice, and each one has to carry the emotion underneath it.

  2. Music

    We use AI generated music. We're not musicians, and making good original music is an enormous amount of work. We could license good music, but it's not worth the hassle for something that plays for 10 to 15 seconds at most. In most anime, music shows up in short segments just to evoke a specific feeling, and outside of the opening and ending songs, few fans pay close attention to it, so it's not worth the investment for us right now.

    Down the line, once we're making a full length film, hiring a composer will make a lot more sense, a longer runtime justifies a longer soundtrack, though from what we've seen, a lot of shows just reuse the opening or ending theme instead.

  3. SFX

    Sound effects (footsteps, impacts, magic, ambient environment sound) are almost entirely AI generated at this stage instead of pulled from a library. We can create exactly the sound we need, and honestly, it's just easier for us.

  4. Room tone for dialogue

    A thin, consistent layer of ambient noise sits under every line of dialogue, so the dialogue doesn't feel empty. We've found ElevenLabs works really well for this.

By the time editing is done, we've probably watched the content 100+ times. We hit export. Then do this over again

That's the whole process, from idea to final export. I've probably skimmed a few stuff, so if you have questions, feel free to reach out. I would love to yap

联系我们 contact @ memedata.com