Coffee Intern

Recreate the reference as a candid office selfie from the uploaded cat's point of view. The uploaded cat sits at a messy corporate desk, one front paw extended toward the camera as if taking the selfie, squinting with a tired unimpressed expression while drinking iced coffee through a straw from a large clear plastic cup in the foreground. Office monitors, keyboard and coworkers blurred behind, fluorescent office light, vertical phone-photo crop, lanyard badge with a tiny photo of the same cat near the lower edge. Preserve the uploaded pet's actual species, breed, coat color, markings, ear shape, eye color, muzzle and facial structure. Do not assume a fixed breed, color, or species from the reference image; adapt the exact gameplay to the uploaded pet. Actively change pose, facial expression, gaze direction, camera angle, wardrobe and styling to match the gameplay reference. Match the gameplay reference very closely in composition, crop, texture, lighting, visible pet angle, props, background, graphic layout and image quality. Vertical 3:4. Keep only the text explicitly requested in the prompt; do not add unrelated logos, watermarks or random words.

Use Template
arrow
Coffee Intern

More From VIVAGO AI

Puppy Stretch AI effects generated image

Puppy Stretch

"Use the uploaded pet image as the exact identity reference. Preserve the pet's recognizable face, fur color, markings, eye color, muzzle shape and ears while transforming it into the requested gameplay style. Create a vertical 3:4 image with this exact effect: realistic humorous viral pet photo, the pet on a warm wooden floor in a yoga/cobra stretch pose, front chest and front paws on the floor, head raised and looking upward with a blissful funny expression. Only the hind legs, hips and lower torso wear snug pale gray yoga pants / leggings; the head, ears, front chest and front paws stay natural and uncovered. Match the reference angle and body exposure: stretched lower body behind, visible fitted gray leggings on the back half, natural upper body in front, clean indoor light. Do not make it a blanket, burrito wrap, full onesie, hoodie, pajamas, or full-body costume. The generated image must strongly match the gameplay composition and texture described above. Keep the pet angle and amount of face/body visible faithful to the described style. If the style is hand-drawn, watercolor, felt, plush, sticker, toy, comic, or package design, reproduce that material and texture clearly. Keep the pet identifiable as the uploaded pet even when stylized. No logos, no watermark, no Doubao AI text."

Cool Commute AI effects generated image

Cool Commute

"Recreate the reference as a black-and-white cool pet portrait. The uploaded dog faces forward wearing a black baseball cap pulled low over the eyes and white wired earbuds hanging down from both ears. Plain gray studio background, monochrome photo, calm serious expression, chest-up crop, soft fur detail, minimal streetwear attitude. Preserve each uploaded pet's actual species, breed, coat color, markings, ear shape, eye color, muzzle and facial structure. Do not assume a fixed breed, color, or species from the reference image; adapt the reference gameplay to the uploaded pet. Actively change pose, facial expression, gaze direction, camera angle and styling to match the gameplay reference. No extra animals beyond the requested uploaded pets. No human face unless explicitly requested; anonymous hands are allowed only when the gameplay needs hands. No brand logos, no watermark, no random readable text. Match the gameplay reference very closely in composition, crop, texture, lighting, pet expression, visible face angle, props, background, and image quality. Vertical 3:4, photorealistic unless the reference is a poster or sticker layout. "

Viral Dance

"Photorealistic film portrait photography, horizontal rectangular frame, fixed composition rule: 【Character】 is strictly centered in frame, the bottom cropping edge of frame cuts exactly at the character's knees, a thin strip of light beige flat ceiling is reserved on the top of frame; Precise Scene Details Background structure: A vertical square cream solid partition pillar right behind 【Character】, two identical dark walnut solid wood full-height open bookshelves symmetrically placed on both sides of pillar, each shelf built with fixed 5 layers; every shelf densely packed with multicolored hardcover books, most books stand vertically, small stacks of horizontally placed thick books on right part of each shelf with mixed red, blue, brown, off-white book spines; Left side of frame: Off-white fabric single armchair at bottom-left with short dark timber legs, one light khaki square linen cushion leans against seat; a floor lamp with thin metal pole + tapered white fabric lampshade stands in the gap between sofa and left bookshelf; Right side of frame: Partial curved edge of matte white round low coffee table is visible behind character’s right hip; Floor design: Full floor covered with oat beige short loop pile carpet, soft subtle natural shadows from bookshelf and armchair on carpet surface; Lighting & Quality Soft diffused front indoor natural lighting, neutral warm white color temperature, no harsh sharp hard shadows, natural layered ambient shadow; 8K ultra-realistic high definition, intricate textures of wood grain, book spine print and fabric weave, minimalist American home study interior, no extra redundant ornaments inside the whole scene."

Seashell Venus Pet AI effects generated image

Seashell Venus Pet

Recreate a mythic Renaissance sea painting with the uploaded pet seated upright on a large open scallop shell rising from green-blue waves. Long flowing golden-red hair-like fur ribbons stream in the wind, pale luminous body, distant coast and soft sky behind. Tempera-like old-master texture, graceful centered composition. The subject remains an animal throughout. No people or human hands. No real character names, celebrity names, brands, logos, watermarks or readable text. Accurately preserve the uploaded pet's species, breed, coat color and markings, ear shape, muzzle and face structure, eye color, body proportions and tail. Keep it instantly recognisable as the same animal; change only pose, scene, costume and art treatment. Match the reference composition closely. Vertical 3:4, crisp detail, the pet is the unmistakable main subject.

Collage Poster AI effects generated image

Collage Poster

"Ultra-realistic vintage cute portrait collage in the late 2000s style, featuring a multi-panel layout that showcases 5 to 6 different poses of the figure from the uploaded image (with unchanged facial features, age and gender) and natural facial retouching with a fresh sheer makeup look: making a peace sign, blowing a pink bubble gum, resting her cheek on one hand while holding a small white camera, standing with one hand on her hip, cuddling a tabby cat, and holding a bouquet of daisies. The girl has long hair with pink-to-purple gradient streaks and a fresh, cute makeup look. Attire: A pastel rainbow gradient cardigan (with light purple/light yellow/light blue stripes), a light purple high-waisted mini skirt, a thin white waist belt, rainbow-striped athletic socks, and white casual sneakers. Accessories: A dopamine colorful Y2K necklace, delicate colorful floral hair clips, and a colorful pendant necklace. Shooting Angles: Mixed perspectives (close-up facial shots, bust shots, full-body shots), captured from the angles of casual natural lifestyle photography. Lighting: Bright and soft studio lighting with a textured translucent sheen, pale shadows, creating a fresh and warm atmosphere. Color Scheme: Macaron soft tones (light purple/light pink/light blue/light yellow) adorned with collage decorations (star/butterfly/heart stickers, sequins), featuring bright low-saturation hues that evoke a vintage cute early 2000s vibe. Layout: A playful scattered arrangement with the effect of vintage magazine clippings, accented with text elements such as ""SO CUTE!"", ""1990S!"", and ""GIRL VIBES""."

Moon Cap AI effects generated image

Moon Cap

Use the uploaded pet image as the only identity reference. Create a vertical 3:4 cute hand-drawn sticker sheet illustration that closely matches the reference style. Strictly preserve the uploaded pet’s recognizable face, fur color, markings, eye color, ear shape, muzzle shape, and overall identity. Show the same pet repeated in 6–8 different cute bedtime poses across the page. The pet should remain clearly recognizable in every pose. Visual style: - soft hand-drawn watercolor-pencil illustration - warm cream paper background - visible soft paper grain - gentle pastel coloring - slightly faded colored-pencil texture - cozy storybook / stationery sticker-sheet feel - soft blue outline around each sticker - white sticker border around each pet drawing - rounded adorable proportions - sweet, gentle, sleepy mood Bedtime theme details: - blue-and-white striped sleep cap with fluffy pom-pom - sleepy and cozy expressions - lying down, peeking, sitting, hugging a pillow, hugging a fish toy, sleeping on paws, rolling over - small blue decorative doodles scattered between stickers - crescent moon - little fish - tiny hearts - stars - small dots - one tiny simple doodle cat face or similar cute doodle in the middle if space allows Composition: - sticker-sheet layout with evenly distributed pet poses - each pose separated clearly like individual stickers - clean balanced arrangement - the page should feel full but not crowded - keep all stickers fully visible inside the frame Important: - keep the pet identifiable as the uploaded pet in every pose - emphasize watercolor pencil texture and hand-drawn softness - not photorealistic - not 3D - not glossy digital art - no logo - no watermark - no random text - no “Doubao AI” text

Fur Crystals AI effects generated image

Fur Crystals

Recreate the luxury editorial portrait. The uploaded pet is shown from chest up in three-quarter view wearing a voluminous blush-pink faux-fur stole and multiple oversized clear crystal necklaces stacked around the neck. Dark gray studio background, cool glamorous light, elegant alert gaze, detailed sparkling gems and fluffy fur. The subject remains an animal throughout. No people or human hands. No real brands, logos, watermarks or readable text. Accurately preserve the uploaded pet's species, breed, coat color and markings, ear shape, muzzle and face structure, eye color and body proportions. Keep it instantly recognisable as the same animal; change only pose, scene, clothing and accessories. Vertical 3:4, photorealistic, crisp detail, with the pet as the unmistakable main subject.

WaveLilt

"8K ultra-high definition realistic photography, mobile candid daily photo style, vertical composition, slightly overhead perspective, camera facing the center of the frame, extreme detail restoration, realistic texture, no filters 【Fixed Scene】: Foreground is a glossy white smooth ceramic sink, half filled with milky white warm water, with slight natural ripples and soft reflections on the water surface, the complete lower body outline of the subject is clearly visible underwater; upper left of the frame is a glossy chrome double-handle faucet with clear highlight reflections on the metal surface, a black overflow hole with a round silver border on the sink wall to the right of the faucet; overall soft and even indoor top lighting, no harsh shadows, clean bright warm white tone There is one animal, extremely tiny and petite, with its overall size exactly one-fifth of the sink, standing normally in the water of the sink, upper body above the water surface, lower body fully submerged in water, all limbs fully visible without occlusion; front paws placed in front of the chest in a relaxed posture, naive and innocent expression, eyes looking directly at the camera. The subject wears a sky-blue disposable non-woven shower cap with mesh breathable top and same-color elastic edge, fully covering head fur without exposure; keep the original appearance and facial features unchanged, consistent with the reference image."

Good Night AI effects generated image

Good Night

Cute emoticon stickers, presented in a realistic photography style. The uploaded pictures show the protagonist (whose race and facial features remain unchanged) curled up and sleeping peacefully, with closed eyes, wearing a cute plush sleep cap, lying on a cute pillow with patterns, presenting a relaxed sleeping posture. The slightly furrowed eyebrows carry a faint nostalgic "saudade" sentiment. The faint melancholy contrasts sharply with the gentle sleepiness. The background is a cold-toned dreamy starry sky, with a soft crescent moon and twinkling stars emitting warm golden light. The nighttime atmosphere is the shining blue and yellow stars and moon with cold dark tones. At the bottom, there is a beautiful handwritten artistic circular warm pale yellow Portuguese text "Good night, saudade!" presented in a smooth hand-drawn cartoon style, with soft ambient light, matte texture, delicate hair details, a dreamy and healing feeling, clear borders, and no messy elements. Brazilian cultural irregular stickers, presented in ultra-high-definition 8K format, are super cute.

Bridal Cruise AI effects generated image

Bridal Cruise

"Create a photorealistic cinematic wedding portrait of the two subjects from the uploaded image, preserving their exact faces, ages, facial features, skin tones, hairstyles, and overall likeness as accurately as possible. Transform the two subjects into a stylish bride and groom posing in and beside a vintage red convertible car after their wedding. The scene should feel like a luxurious post-wedding street fashion editorial, romantic, confident, glamorous, and slightly rebellious. The groom sits on the left side inside the open-top vintage convertible car, wearing a black tuxedo and black sunglasses. He leans back casually in the driver’s seat or front seat, one arm resting naturally on the car door or steering wheel, with a calm, cool, confident expression. His posture should feel relaxed, elegant, and masculine, like a high-end fashion wedding photo. The bride sits or reclines on the right side beside the groom, leaning back against the red leather car seat with a confident and glamorous pose. She wears white cat-eye sunglasses and looks toward the camera with a cool, romantic, slightly playful expression. Her body language should feel relaxed, stylish, luxurious, and confident, as if enjoying a glamorous wedding getaway. Bride outfit: the bride wears a dramatic white wedding gown inspired by the reference image. The dress has a strapless or sweetheart neckline, sculptural ruffle details around the bodice or sleeves, layered soft tulle, a voluminous flowing skirt, and a high slit that subtly reveals one leg. The dress should look fashionable, romantic, luxurious, and editorial. The fabric should have realistic soft tulle texture, natural folds, and elegant volume. The bride wears white strappy high-heeled sandals, subtle bridal earrings, and a delicate pearl or diamond necklace. A soft white bridal veil may flow naturally behind her or blend into the dress, without covering her face. Groom outfit: the groom wears a classic black tuxedo, fitted and elegant, with a clean white dress shirt, black bow tie, black dress shoes, and stylish black sunglasses. The tuxedo should look sharp, formal, minimal, and high-end. Do not use a casual suit, colorful suit, necktie, or open-collar shirt. Car and background: a vintage red convertible car with a classic luxury feel, open roof, polished red exterior, red or dark leather interior, realistic steering wheel, chrome details, and an elegant retro design. The couple is parked on a quiet European-style city street or upscale residential boulevard. The background includes soft greenery, iron gates, warm stone buildings, palm trees or street trees, and sunlit urban details, softly blurred with shallow depth of field. No other people should appear in the image. Lighting and color: warm daylight wedding editorial photography, natural sunlight, soft shadows, realistic skin texture, subtle highlights on sunglasses, tuxedo, wedding dress, car paint, and chrome details. Use an elegant low-to-medium saturation color palette with warm beige tones, soft ivory whites, deep black tuxedo contrast, muted red car tones, gentle green background, and refined cinematic color grading. The image should feel expensive, fashionable, romantic, and timeless, avoiding overly vivid or neon colors. Composition: vertical portrait, medium-full or full-body framing. The couple and the red convertible car should be the central focus. The groom is on the left inside the car, the bride is on the right, leaning elegantly with her white wedding dress flowing across the car seat and lower part of the frame. The bride’s legs and white strappy heels may be visible naturally, but the pose should remain elegant and tasteful. Both faces should be sharp, recognizable, and clearly visible behind sunglasses. The overall mood should feel like a glamorous “just married getaway” moment. Ultra realistic photography, high-end wedding editorial style, luxury fashion photography, realistic car interior, realistic fabric texture, realistic tulle layers, realistic sunglasses reflections, natural facial details, soft film grain, professional camera look, high detail, cinematic muted color grading, stylish post-wedding atmosphere. Negative prompt: extra people, background people, crowd, wedding guests, driver, duplicated bride, duplicated groom, distorted faces, changed identity, inaccurate likeness, deformed hands, extra fingers, missing fingers, fused fingers, broken fingers, unnatural arms, stiff pose, awkward sitting pose, cheap modern car, closed roof car, wrong car color, missing sunglasses, sunglasses covering the entire face, missing wedding dress, wrong wedding dress, plain casual dress, colorful wedding dress, casual groom outfit, necktie instead of bow tie, open-collar shirt, plastic skin, over-smoothed face, fake CGI look, text, logo, watermark, title text, low resolution, blurry face, oversaturated colors, vivid bright colors, neon colors, harsh contrast, black and white, monochrome, grayscale."

Load more

Next-Gen Multi-Model AI Video Architecture

Vivago AI isn't just one engine—it’s a unified hub for the world’s most advanced video AI. Whether you need cinematic realism or high-speed social content, we provide the right model for your creative vision.

Free Generate

Beauty and Dolphins

Vacation Time

Stellar Tear

Fish Tank Supervisor

Cinematic Quality & Precision Control

Enables 4K resolution with multi-lens motion control, generating delicate scene via text prompts for customized cinematography.​

TRY NOW

Dynamic AV Sync

Auto-generates original audio to avoid copyright issues. Build 3D immersive environments through layered sound design automatically.

TRY NOW

OpenAI Sora 2

​Advanced visual storytelling with unparalleled physics and consistency.

TRY NOW

Kling v2.6 Pro

Industry-leading cinematic image animation and motion control.

TRY NOW

Google Veo 3 & 3.1

Ultra-fast generation with enhanced realism for creative workflows.

TRY NOW

Vivago AI 2.0

Our proprietary model optimized for efficiency, speed, and cost-effective generation.

TRY NOW

Users' Voice

We listen carefully to the opinions of every user.
Free Generate
Contact Us
I tried the Lip Sync feature inside Vivago.ai’s AI Video Generator for my educational podcast, and the results were stunning! The avatar's lip movements perfectly matched my audio recording, creating a professional AI-generated video without complex editing. Compared with tools like OpenAI Sora 2 and Google Veo 3.1, Vivago Image-to-Video delivers fast, studio-quality results online. It saved me hours of post-production work.
ElenaM (Spain)
Vivago’s Image-to-Video AI transformed my marketing workflow. I uploaded a product image and described the launch scene in text, and it generated a 10-second cinematic AI video with background music and dynamic visuals. The output quality rivals Kling v2.6 Pro and Google Veo 3 Fast. It’s now my go-to AI video generator for social media ads and product campaigns.
KenjiT (Japan)
As a digital artist, I use Vivago.ai 2.0 daily for Image-to-Image and AI Image-to-Video creation. The e-book covers and animated visuals I generate for clients look cinematic and professional. Unlike many standalone AI tools, Vivago integrates multiple leading models into one platform, making it easier to create copyright-safe AI images and videos for publishing.
ChenL (China)
I tried the Lip Sync feature inside Vivago.ai’s AI Video Generator for my educational podcast, and the results were stunning! The avatar's lip movements perfectly matched my audio recording, creating a professional AI-generated video without complex editing. Compared with tools like OpenAI Sora 2 and Google Veo 3.1, Vivago Image-to-Video delivers fast, studio-quality results online. It saved me hours of post-production work.
ElenaM (Spain)
Vivago’s Image-to-Video AI transformed my marketing workflow. I uploaded a product image and described the launch scene in text, and it generated a 10-second cinematic AI video with background music and dynamic visuals. The output quality rivals Kling v2.6 Pro and Google Veo 3 Fast. It’s now my go-to AI video generator for social media ads and product campaigns.
KenjiT (Japan)
As a digital artist, I use Vivago.ai 2.0 daily for Image-to-Image and AI Image-to-Video creation. The e-book covers and animated visuals I generate for clients look cinematic and professional. Unlike many standalone AI tools, Vivago integrates multiple leading models into one platform, making it easier to create copyright-safe AI images and videos for publishing.
ChenL (China)
I absolutely love Vivago’s AI Image-to-Video Generator. As a travel blogger, static images often fail to capture real atmosphere, but Vivago helps me turn photos into vivid cinematic AI videos with motion effects. It feels comparable to OpenAI Sora 2 and Google Veo 3.1, but more accessible and faster for creators who need high-quality AI videos online.
LiamK (Australia)
I tried the Lip Sync feature inside Vivago.ai’s AI Video Generator for my educational podcast, and the results were stunning! The avatar's lip movements perfectly matched my audio recording, creating a professional AI-generated video without complex editing. Compared with tools like OpenAI Sora 2 and Google Veo 3.1, Vivago Image-to-Video delivers fast, studio-quality results online. It saved me hours of post-production work.
ElenaM (Spain)
Vivago’s Image-to-Video AI transformed my marketing workflow. I uploaded a product image and described the launch scene in text, and it generated a 10-second cinematic AI video with background music and dynamic visuals. The output quality rivals Kling v2.6 Pro and Google Veo 3 Fast. It’s now my go-to AI video generator for social media ads and product campaigns.
KenjiT (Japan)
As a digital artist, I use Vivago.ai 2.0 daily for Image-to-Image and AI Image-to-Video creation. The e-book covers and animated visuals I generate for clients look cinematic and professional. Unlike many standalone AI tools, Vivago integrates multiple leading models into one platform, making it easier to create copyright-safe AI images and videos for publishing.
ChenL (China)
I absolutely love Vivago’s AI Image-to-Video Generator. As a travel blogger, static images often fail to capture real atmosphere, but Vivago helps me turn photos into vivid cinematic AI videos with motion effects. It feels comparable to OpenAI Sora 2 and Google Veo 3.1, but more accessible and faster for creators who need high-quality AI videos online.
LiamK (Australia)
Using Vivago.ai’s Image-to-Video AI has greatly enhanced my classroom teaching. I transform textbook notes into historical AI videos with cinematic filters and dynamic animations. Compared with tools like Kling v2.6 Pro and Google Veo 3 Fast, Vivago offers faster generation and easier parameter control for educators who need reliable AI video creation.
RajivG (India)
I frequently create AI videos on Vivago and publish them on TikTok and YouTube Shorts. The AI video templates and trending content ideas help me produce viral-ready clips quickly. With Vivago’s integrated models—including advanced video engines similar to OpenAI Sora 2—I can generate anime-style and cinematic social media videos that drive high engagement.
MarieJ (Spain)
What attracts me most about Vivago.ai is not only the powerful AI Video Generator but also the active AIGC creator community. It combines AI Image-to-Video, Text-to-Video, and leading model integrations like Google Veo 3.1 into one creative platform.
TomW (India)
At first, I was hesitant about using AI video tools. But after trying Vivago Image-to-Video, I realized how easy it is to create professional AI-generated videos online. I just upload an image, add a short prompt, and adjust a few settings. The results are cinematic and copyright-safe, which is essential for commercial projects.
HectorC (Mexico)
Using Vivago.ai’s Image-to-Video AI has greatly enhanced my classroom teaching. I transform textbook notes into historical AI videos with cinematic filters and dynamic animations. Compared with tools like Kling v2.6 Pro and Google Veo 3 Fast, Vivago offers faster generation and easier parameter control for educators who need reliable AI video creation.
RajivG (India)
I frequently create AI videos on Vivago and publish them on TikTok and YouTube Shorts. The AI video templates and trending content ideas help me produce viral-ready clips quickly. With Vivago’s integrated models—including advanced video engines similar to OpenAI Sora 2—I can generate anime-style and cinematic social media videos that drive high engagement.
MarieJ (Spain)
What attracts me most about Vivago.ai is not only the powerful AI Video Generator but also the active AIGC creator community. It combines AI Image-to-Video, Text-to-Video, and leading model integrations like Google Veo 3.1 into one creative platform.
TomW (India)
At first, I was hesitant about using AI video tools. But after trying Vivago Image-to-Video, I realized how easy it is to create professional AI-generated videos online. I just upload an image, add a short prompt, and adjust a few settings. The results are cinematic and copyright-safe, which is essential for commercial projects.
HectorC (Mexico)