Início/Explorar/Wan 2.1 Video Models/wavespeed-ai/wan-2.1/text-to-image
text-to-image

text-to-image

Wan 2.1 Text-To-Image

wavespeed-ai/wan-2.1/text-to-image

Wan 2.1 Text-to-Image delivers ultra-realistic photographic images by adapting the Wan 2.1 video model for SOTA visual fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Hint: You can drag and drop a file or click to upload

width
height
If enabled, the output will be encoded into a BASE64 string instead of a URL. This property is only available through the API.
If set to true, the function will wait for the result to be generated and uploaded before returning the response. It allows you to get the result directly in the response. This property is only available through the API.

Idle

Envision an ethereal and highly decorative portrait of an androgynous Elven Monarch, seated upon a throne carved from living, iridescent wood within a moonlit glade. Their form is framed by the elegant, sinuous 'whiplash' curves characteristic of Art Nouveau. Long, flowing silver hair, intricately braided with glowing flora and pearls, cascades around them. They are adorned in gossamer robes of silk and moonlight, featuring delicate, repeating patterns of lilies and dragonfly wings. Their expression is serene and ancient, with eyes holding a gentle, knowing light. One hand elegantly gestures, causing magical, opalescent petals to swirl in the air. The background is a flat, decorative tapestry of intertwined vines, stylized trees, and celestial motifs, rendered in a soft palette of muted lavenders, sage greens, and creamy golds, with intricate gold leaf detailing. The entire composition is a harmonious symphony of organic forms and graceful lines, celebrating beauty, nature, and magic with the quintessential elegance of the Art Nouveau masters.

Sua solicitação custará $0.02 por execução.

Por $1 você pode executar este modelo aproximadamente 50 vezes.

Mais uma coisa::

ExemplosVer todos

Envision an ethereal and highly decorative portrait of an androgynous Elven Monarch, seated upon a throne carved from living, iridescent wood within a moonlit glade. Their form is framed by the elegant, sinuous 'whiplash' curves characteristic of Art Nouveau. Long, flowing silver hair, intricately braided with glowing flora and pearls, cascades around them. They are adorned in gossamer robes of silk and moonlight, featuring delicate, repeating patterns of lilies and dragonfly wings. Their expression is serene and ancient, with eyes holding a gentle, knowing light. One hand elegantly gestures, causing magical, opalescent petals to swirl in the air. The background is a flat, decorative tapestry of intertwined vines, stylized trees, and celestial motifs, rendered in a soft palette of muted lavenders, sage greens, and creamy golds, with intricate gold leaf detailing. The entire composition is a harmonious symphony of organic forms and graceful lines, celebrating beauty, nature, and magic with the quintessential elegance of the Art Nouveau masters.
A candid, spontaneous selfie of a stunning young Black woman standing on her balcony at golden hour, dressed in a cropped white tee and wearing gold hoop earrings that catch the warm sunlight. Her curly hair frames her glowing skin naturally, illuminated by the soft last rays of the sun. The blurred city rooftops behind her fade gently into an orange-toned twilight sky. The image features authentic skin texture with subtle highlights and natural shadows, casual framing with a slight tilt capturing the intimate and effortless moment. The overall lighting and ambience reflect typical warm, natural light of an iPhone photo, making the scene feel genuine and elegantly powerful.
A platinum bob slips loose from a claw-clip, framing a woman’s face with frosted lips and a smudge of sunlit freckles. Her hand presses lightly to her cheek, stacked with chunky silver rings that glint like molten metal, bending in fluid, bold shapes. Black wrap-around sunglasses catch the faint reflection of city glass and the phone’s screen, while cool overcast light throws soft shadows on skin with pores and the fuzz of a wool rib cuff nearby. The shot tilts just off-center, fall-off blur wrapping around her cheek and shading, capturing the casual disarray of an artful moment cropped tight—close-up captured on Iphone, hand-face jewelry focus
A side-view photo of a cat walking gracefully along a narrow balcony railing at night. The background reveals a softly blurred city skyline glowing with lights—windows, streetlamps, and distant cars forming a bokeh effect. The cat's fur catches subtle reflections from the urban glow, and its tail balances high as it steps with precision. Cinematic night lighting, shallow depth of field, high-resolution photograph.
Intense medieval battle scene with knights clashing in close combat, swords swinging, and shields raised. The image is filled with dynamic motion blur — blurred swords, flying debris, and rushing figures convey the chaos of the fight. Dust rises from the ground, kicked up by charging horses and running soldiers. Armor glints in the sunlight, partially obscured by blur and dirt. The composition captures the raw energy of the battlefield, with blurred foreground action and slightly sharper figures in mid-ground. Gritty, cinematic lighting, overcast sky, and a muted, earth-toned color palette. Shot with a wide lens, slightly tilted, as if captured in the middle of battle.
Close-up of a woman’s hand fully submerged in clear ocean water, elegantly holding a bright yellow lemon. Her nails are clean and simple. Around the hand, small tropical fish swim gracefully, with strands of seaweed and soft marine plants drifting nearby. Sunlight filters through the water surface above, casting dappled light and gentle caustics on the skin and surrounding sea life. The skin has hyperrealistic detail — soft, luxurious, and radiant. The water is turquoise-green, slightly hazy with floating particles, creating a dreamy, cinematic underwater atmosphere. The scene feels elegant, epic, and editorial — a surreal blend of luxury and nature.
High‑quality photo. Black woman with a big afro leans on a mustard‑yellow 1970s convertible at a retro gas station. She wears high‑waisted orange flared pants and a tucked‑in paisley shirt that shows her waist. Large gold hoop earrings move slightly in the desert breeze. Warm late‑afternoon light falls on the glossy car and her face. Dusty asphalt, vintage pumps, and faded signs appear sharp. Pastel dusk sky fills the background. Shot eye‑level on a 50 mm lens, Kodak Gold 100 film with visible grain and a touch of lens flare. Late‑70s / early‑80s cinematic look.
A european woman with short, tousled hair leans in close to a man, her eyes gently closed. She wears a glowing red sweater; he has a jeans jacket with a bright collar. Golden ambient light softens their skin tones. Their faces are calm and close, framed tightly in an intimate, cinematic shot. The background is a gentle blur of city motion, adding contrast to their stillness. Subtle film grain evokes a timeless, romantic feel.
Close-up, top-down view of a teenage girl lying on a vibrant picnic blanket spread out on the green lawn of a sunny backyard garden. She’s laughing joyfully as a playful puppy stands beside her head, licking her ear. Her eyes are closed from laughter, and her expression is full of pure delight. The blanket is colorful — with bright patterns like florals or stripes — contrasting against the lush grass. Around them, soft natural sunlight filters through tree leaves, casting dappled shadows and warm highlights across her face, hair, and the puppy’s fur. Flowers, garden plants, or scattered toys add playful detail in the background. A lively, heartwarming outdoor moment filled with summer energy and color.
Close-up of a woman in ancient costume, with soft light falling on her skin, outlining delicate contours.
A close-up portrait of a cheerful Scandinavian man with bright blue eyes, enjoying an outdoor coffee in a quaint European town square bathed in warm morning light. The scene is bright and inviting, with crisp focus on his happy expression.

README

wan-2.1/text-to-image

Wan 2.1 is part of the Wan 2.1 foundation model suite, an advanced AI system developed to redefine video and image generation. This model focuses on text-to-image synthesis — transforming detailed written prompts into vivid, high-resolution visuals with cinematic precision.

🌟 Key Features

  • 🎨 SOTA Image Quality Built on Wan 2.1’s next-generation video foundation, this model produces exceptional still-frame quality with realistic lighting, texture, and depth.

  • 🧠 Multilingual Understanding Supports both Chinese and English prompts, ensuring accurate and context-rich image generation across languages.

  • ⚙️ Fine Control with Parameters Adjustable inputs such as strength, width, and height provide creators with direct control over composition and style.

  • 🪄 Powerful Visual Consistency Based on Wan-VAE, enabling coherent detail, color fidelity, and stylistic alignment across resolutions.

  • 💰 Lightweight and Efficient High-quality generation at a base cost of just $0.02 per image, ideal for scalable creative workflows.

⚙️ Parameters

ParameterDescription
prompt*Text description of the image to be generated (supports CN/EN).
image(Optional) Upload a reference image for guided generation.
strengthControls how strongly the image follows the prompt or reference (0–1).
size (width / height)Define custom output resolution; max recommended ratio 2:1.
seedFix for reproducibility or randomize for variation.
output_formatChoose from jpeg, png, or webp.

💡 Example Prompt

Envision an ethereal and highly decorative portrait of an androgynous Elven Monarch, seated upon a throne carved from living iridescent wood within a moonlit glade. Intricate Art Nouveau details, luminous textures, soft-focus background, cinematic lighting.

💰 Pricing

MetricPrice
Per image generated$0.02 / image

🎯 Use Cases

  • Concept Art & Illustration — Generate fantasy, sci-fi, or cinematic character art.
  • Visual Design & Branding — Create unique imagery for marketing, web, or product visuals.
  • Research & Visualization — Produce clear, detailed concept visuals from descriptive text.
  • Previsualization — Generate cinematic stills for film, animation, or game design workflows.