Poster and Key Visual Design From AI Stills
How a generated frame becomes a campaign key visual, what resolution and layout demand, and why the poster should be designed before the film.
The key visual is the image a campaign is remembered by. It appears on the poster, the backdrop, the banner, the social card, the invitation and the endframe of the film, and it carries the campaign's identity across every format. Because it is a still, it is often treated as an output of the film rather than an input to it, which is the wrong order and produces weaker work in both.
Designing the key visual first has a practical benefit beyond the poster. It settles the world, the light, the palette and the treatment of the subject as a single agreed image, which then governs the film. This is the same discipline as the stills first workflow, applied at the level of the campaign rather than the shot, and it means the film and the printed material share a visual language by construction rather than by effort.
Composition for a key visual differs from composition for a film frame in one important way: it must survive having text placed on it. That means deliberate negative space, usually a third of the frame, positioned where the headline and logo will sit, with tonal values in that region controlled so type remains legible. A beautiful frame with the subject centred and detail everywhere is a poor key visual, because every layout option damages it.
It must also survive extreme reformatting. The same image has to work as a portrait poster, a landscape banner, a square social card and a wide event backdrop, and these are not crops of one another. The reliable method is to generate or render the scene wider and taller than any single format requires, keeping the subject and the negative space positioned so that each crop remains balanced. Deciding the crops in advance costs nothing and prevents rebuilding the image per format.
Resolution is the constraint that catches generated key visuals out. A frame that is adequate for a film is far below what a printed poster or a large format backdrop requires. Generating at the highest available resolution and then upscaling carefully is the usual route, and it has limits: roughly a doubling of linear dimensions holds up, beyond which edges acquire the over sharpened, plasticky quality that is obvious in print. For very large output the safer approach is to build the key elements in 3D and composite.
Brand elements must be exact and therefore should not be generated. The logo, the typography, the product geometry and any packaging copy are placed on top of the generated environment as accurate artwork. Kirk and Givi (2025) found that perceptions of authenticity shape consumer responses to AI generated marketing communications, and a key visual is the most scrutinised image in a campaign, seen at size, in print, for weeks. A subtly wrong logo on a poster is a lasting problem rather than a fleeting one.
Colour management matters more here than anywhere else in the campaign because the image will be printed. A frame that looks correct on a screen may shift substantially in CMYK, particularly in saturated blues, greens and oranges, which are exactly the colours generative models produce enthusiastically. Converting early, proofing on the actual stock, and adjusting the source rather than the print is the sequence that avoids a poster that does not match the digital campaign.
The emotional work of the palette should be deliberate rather than inherited from whatever the model produced. Jonauskaite et al. (2020) documented consistent patterns of emotion associations with colours, and a key visual that will represent a campaign for months is worth choosing rather than accepting. Limiting the image to three or four values with defined roles produces a poster that reads at distance, which is the actual test of the format.
Kong and Lou (2026) found that visual appeal and visual congruence play distinct roles alongside social proof in how advertising is processed. For a key visual this is a useful pair of criteria: appeal determines whether the image is noticed, and congruence between the image and the product determines whether the attention converts into anything. A striking image unrelated to what is being sold performs the first job and fails the second.
The practical workflow is to generate three or four genuinely different key visual directions as stills, present them at actual layout size with the headline and logo in place rather than as bare images, choose one, then build the campaign and the film from it. Presenting bare frames without the layout is the most common cause of a key visual approved in isolation and abandoned once the text arrives.
References
Kirk, C. P., & Givi, J. (2025). The AI-authorship effect: Understanding authenticity, moral disgust, and consumer responses to AI-generated marketing communications. Journal of Business Research, 186, Article 114984. https://doi.org/10.1016/j.jbusres.2024.114984
Jonauskaite, D., Parraga, C. A., Quiblier, M., & Mohr, C. (2020). Feeling blue or seeing red? Similar patterns of emotion associations with colour patches and colour terms. i-Perception, 11(1), Article 2041669520902484. https://doi.org/10.1177/2041669520902484
Kong, J., & Lou, C. (2026). Beyond persuasion knowledge: Examining the roles of visual appeal, visual congruence, and social proof in influencer advertising. Journal of Retailing and Consumer Services, 88, Article 104502. https://doi.org/10.1016/j.jretconser.2025.104502