Generative AI47
How AI makes images, video and music, and how to use it. Diffusion, GANs, prompt engineering, negative prompts, seeds, upscaling: the terms you meet the moment you open an image generator.
Turning a few photos or a line of text into a 3D shape
AI & CopyrightThe open questions of rights around AI training and AI output
AI ArtPictures made with generative tools, and the debates around them
AI WatermarkAn invisible mark quietly embedded in what AI generates
AutoencoderA structure that learns by shrinking data down and rebuilding it
Background RemovalErasing a photo's background and keeping just the subject
ColorizationAdding fitting color to a black-and-white picture
ConsistencyKeeping the same subject looking the same across every shot
ControlNetA device that pins an image's shape down with a rough sketch
DeepfakeFake video or audio where AI copies a real face or voice
DenoisingRemoving only the flecks mixed in, leaving the original untouched
Diffusion ModelPeeling away blur, layer by layer, until a picture appears
GANGenerative Adversarial NetworkA maker and a checker compete, and both get better
Gaussian SplattingA 3D scene held as millions of scattered color blobs
Generative AIAI that invents a result that never existed before
Guidance ScaleThe dial that decides how literally a request gets followed
Image CaptioningLooking at a picture and writing what's in it as a sentence
Image-to-ImageBuilding a new picture starting from a given one
InpaintingErasing part of a picture and filling it back in
InterpolationFilling the space between two points with several in-between steps
Latent SpaceThe place where compressed values end up sitting
Lip SyncFitting mouth shapes to the beat of a sound
LoRALeaving the main body alone and training only a small add-on
Mode CollapseWhen the generator settles on one winning output and just repeats it
Motion CaptureTurning body movement into a record of point locations
Music GenerationTurning a few lines of request into a finished piece of music
Negative PromptThe field for writing what shouldn't appear in the picture
NeRFNeural Radiance FieldA 3D scene that answers with light for any spot and direction
NoiseThe blur mixed in — and also where generation starts
OutpaintingWidening a picture's frame and drawing the new edges to match
Prompt AdherenceThe yardstick for whether a result matches what was requested
Prompt EngineeringShaping an instruction so the result matches what you want
ResolutionThe number of dots a picture is made of
Sampling StepsThe number of passes it takes to turn blur into a picture
SeedThe starting number that reproduces the same picture again
Style TransferKeeping a picture's content and swapping only its surface look
Synthetic DataPractice data made up in place of the real thing
Text GenerationBuilding text by looking ahead and adding the next piece
Text-to-ImageMaking a picture out of what's written down
Text-to-VideoTurning text or a photo into a moving scene
TTSText-to-SpeechThe technology that reads text out loud in a voice
U-NetShrinks an image down, then grows it back while recovering position
UpscalingEnlarging a picture by inventing the detail it never had
VariationSlightly different results pulled from the same request
Variational AutoencoderAn autoencoder that remembers things as a range, not a point
VocoderThe part that turns a sound blueprint into an actual waveform
World ModelA model of how the world works that you can run before acting