Generative media

Generating one good image is a prompt. Generating four hundred that are all recognisably yours, and none of which embarrass you, is a pipeline.

At any real volume the interesting problem is consistency, not quality. A model will happily produce four hundred excellent images that share no palette, no lighting logic and no sense of being from the same company. Holding a look steady takes deliberate machinery — reference images, fixed seeds where reproducibility matters, constrained palettes, and a post-process that normalises what the model will not. Without it you get a folder of individually good pictures that cannot be used together.

Nothing generated should reach a customer unreviewed. Models produce text inside images that is subtly wrong, hands and products with the wrong number of parts, and occasionally something that resembles a mark you do not own. A queue where a person approves or rejects before publication is not a lack of confidence in the technology; it is the same step any studio has always had between a draft and a release, and it is what makes the output usable commercially.

Voice deserves its own caution, because cloning a voice is a decision with legal and ethical weight rather than a technical feature. Consent needs to be documented, the permitted uses need to be bounded, and the recordings and the resulting model need to be treated as sensitive material. Building this in from the start is straightforward; retrofitting it after a voice has been used in a campaign is not.

How we work

  • Consistency is engineered with references, constrained palettes and post-processing — not hoped for from prompt wording.
  • A human approves before anything is published. Generated media reaching a customer unreviewed is how brands acquire an incident.
  • What was generated, from which prompt and by which model, is recorded — so provenance is answerable later.

What this includes

Pick what you need and send it over.

Related