Input: menu copy, brew method, and cup photo
Output: a six-second pour or steam clip
Aspect: 9:16 and 1:1
Build log
1. Set the rule
I opened the menu photos, recipe steps, and one hero item, locked the real dish shape, portion, and ingredients, and wrote down what could not change.
2. Run one minimum version
I made one small a six-second pour or steam clip first. The crema looked fake, so I used a real reference and reduced the foam.
3. Build the focused variant
I reused the same reference and changed only one pour and one crema close-up.
4. Review and export
I checked dish shape, portion, texture, and no invented ingredients at full size, then exported the final 1:1, 3:4, and 9:16.
Complete execution record
Task boundary and source gate
Treat AI coffee product video as a deliverable, not a definition. The job receives menu copy, brew method, and cup photo and owes a six-second pour or steam clip in 9:16 and 1:1. Write the rejection reason first, then turn it into the first hard constraint.
Require the plated reference, portion size, ingredient list, garnish, table surface, kitchen constraint, and the approved menu wording. Do not invent hygiene claims, health effects, or a preparation step the kitchen does not follow. Recipe showmanship must not replace safe handling.
The imgmov route
Upload the plated dish, hero ingredient, and recipe order. Use cheap generation to prove steam, pour, cut, plating, or texture. Menu facts, price, allergen wording, and safety notes go into the external editor.
FDA separates safe handling into clean, separate, cook, and chill, and says color and texture are unreliable safety indicators. A recipe clip may show process, but safety wording must come from the approved source.
Lock the variables that cannot drift
Lock portion size, hero ingredient color, garnish, plate shape, sauce state, table surface, and any food-safety step. A clip cannot increase the portion.
Put the locked fields at the top of the prompt, not at the end: The practical move is to lock the real dish shape, portion, and ingredients and change only one pour and one crema close-up. Save that prompt with the reference, ratio, and model so the next run starts from a decision instead of a guess.
Step-by-step build path
1) Set the rule: I opened the menu photos, recipe steps, and one hero item, locked the real dish shape, portion, and ingredients, and wrote down what could not change. Leave one checkable artifact from this step; do not start the next until it exists. 2) Run one minimum version: I made one small a six-second pour or steam clip first. The crema looked fake, so I used a real reference and reduced the foam. Leave one checkable artifact from this step; do not start the next until it exists. 3) Build the focused variant: I reused the same reference and changed only one pour and one crema close-up. Leave one checkable artifact from this step; do not start the next until it exists. 4) Review and export: I checked dish shape, portion, texture, and no invented ingredients at full size, then exported the final 1:1, 3:4, and 9:16. Leave one checkable artifact from this step; do not start the next until it exists.
Change one named variable per rerun: prompt, reference, camera, duration, model, ratio, or export crop. If two variables change together, a better output cannot be reused because nobody knows which fix worked.
Review gates and evidence
The hero ingredient must be identifiable, texture must not become plastic, hands must not melt, and menu captions must match what is actually served.
Keep source, approved wording, rejected version, correction, credit cost, and final crop next to the asset. The record exists so the next reviewer can reproduce the decision without a chat thread.
Failure diagnosis and retry ladder
If texture becomes plastic, reduce motion and use a macro still. If sauce flows forever, shorten the pour. If garnish changes, return to the plated reference.
Do not upgrade on hope. A 5-second Agnes Video probe is 10 credits. If the concept passes, Seedance 2.5 costs 90 credits at 480p or 195 at 720p for 5 seconds; Kling is 195 without audio or 245 with audio. Veo enters only when a 4/6/8-second cinematic bucket is genuinely worth 80/125/165 credits. Prove hook, subject, and rhythm first, then pay to clean up motion that already worked.
Versioning and handoff
Menu boards, delivery apps, and short video get separate safe-area checks because price and spicy-level text crop differently.
The folder is boring on purpose: source, approved reference, generation settings, caption file, platform cuts, QA screenshots, and one correction line. A reusable asset is boring in the right way.
Why this is not a generic answer
The imgmov advantage is the chain: Asset Library preserves the subject reference, Canvas fixes shot order and first/last frames, the workspace routes Agnes to Seedance, Kling, or Veo only after proof, and your external editor adds captions and CTA without regenerating the media.
Do not invent hygiene claims, health effects, or a preparation step the kitchen does not follow. Recipe showmanship must not replace safe handling. If the request only asks what ai coffee product video means, a search page is faster. This page is useful when someone must deliver a six-second pour or steam clip under real constraints.
Frequently asked questions
What is the one rule I keep repeating?
The practical move is to lock the real dish shape, portion, and ingredients and change only one pour and one crema close-up.
The workflow is specific about its stopping point: it starts with menu copy, brew method, and cup photo and stops at a six-second pour or steam clip. In imgmov, upload the source, lock it as a reference, set 9:16 and 1:1, run one proof, and save the passing settings. If a reviewer cannot tell which version is approved, the workflow has failed even when the media looks good.
What do I check before export?
Portion, ingredients, texture, text, and final crop.
Do not upgrade on hope. A 5-second Agnes Video probe is 10 credits. If the concept passes, Seedance 2.5 costs 90 credits at 480p or 195 at 720p for 5 seconds; Kling is 195 without audio or 245 with audio. Veo enters only when a 4/6/8-second cinematic bucket is genuinely worth 80/125/165 credits. Prove hook, subject, and rhythm first, then pay to clean up motion that already worked.
What makes this different from a generic generator?
I start from menu copy, brew method, and cup photo and keep a six-second pour or steam clip consistent, instead of inventing a new scene each run.
The check is not “does it look AI-nice?” It is: The hero ingredient must be identifiable, texture must not become plastic, hands must not melt, and menu captions must match what is actually served. Then the source, approved copy, rejected version, correction, and final crop stay in the same handoff folder.
What should I do if the first render drifts?
Cut the motion, lock the reference, and rerun one small version before batching.
If texture becomes plastic, reduce motion and use a macro still. If sauce flows forever, shorten the pour. If garnish changes, return to the plated reference. Change one named variable, keep the old version, and record the credit cost. If the same failure repeats, fix the reference or scope rather than asking the prompt for forgiveness.