LLMs & Generative AI · 2.4

🎨 Generative AI Beyond Text

Images, audio, video — and the ethics that come with them⏱ ~2 min

The same 'learn patterns, then generate new examples' idea works far beyond text. Generative AI can now produce images, music, voices, and video from a text description — with huge creative potential and serious ethical questions.

The Main Types of Generative AI

TypeWhat It DoesExample Tools
TextWrites, summarizes, translates, codesChatGPT, Claude, Gemini
ImageCreates pictures from text descriptionsDALL·E, Midjourney, Stable Diffusion
Audio / VoiceGenerates speech, music, or clones voicesText-to-speech engines, music generators
VideoGenerates or edits video clips from promptsEmerging tools like Sora and others
CodeWrites and explains softwareGitHub Copilot, and coding-focused LLMs

How Image Generators Work (Briefly)

Most image generators use a technique called diffusion. During training, the model learns to remove noise from images. To generate, it starts with pure random noise and repeatedly 'denoises' it, guided by your text prompt, until a coherent image emerges. It's like sculpting a picture out of static.

The Ethics You Have to Think About

⚠ WarningGenerative AI raises real, unsettled questions. Training data often includes copyrighted work scraped from the internet — the legal status is being fought over in courts right now. Generated art can imitate a living artist's style. Voice cloning enables fraud. Deepfakes enable harassment and disinformation. Being a responsible user means thinking about consent, attribution, and harm — not just what's technically possible.
🔒 SecurityAs a creator: check the terms of the tool you use, don't pass off AI work as human when it matters, be transparent when content is AI-generated, and never use voice/face generation to impersonate a real person without consent. These aren't just ethics — increasingly they're the law.
🧠Quick Checkfirst try = +5 XP

Most AI image generators create pictures by…

0 XP🔥 0 days