Midjourney
Midjourney is one of the mainstream AI Image and Design platforms, supporting text-to-image, style reference, parametric control and web page creation workflow, and is widely used in advertising creativity, concept design, illustration and visual exploration.
Midjourney
Midjourney’s core parameters and statistics
Brief review in one sentence: Midjourney is not a "universal AI design platform", but a set of visual generation engines with aesthetic stability as the core - it breaks down high-quality image generation into controllable parameter combinations, allowing designers and creative teams to control the direction of AI output like adjusting camera parameters.
Positioning boundaries: Midjourney officially defines itself as a "community-funded research laboratory" with a team of about 60 people and is self-profit rather than venture capital driven. This background directly affects its product rhythm - it does not pursue full coverage of enterprise-level functions, but continues to dig deeper into visual quality and creative experience. People who are suitable for using it know very well: this is not a material management system, nor a collaborative design platform, it is a highly aesthetic "visual front-end engine".
| Dimensions | Key facts |
|---|---|
| Product positioning | High-quality AI image generation platform for creative and design teams |
| Core Competencies | Text to Image, Image Reference, Style Reference, Parametric Control, Variation and Amplification |
| Latest mainline model | V8.1 Alpha (2026-04-14, set as default model) |
| Main entrance | Midjourney Web (alpha.midjourney.com), Discord Bot |
| Parameter control system | --ar (aspect ratio), --stylize (stylization intensity), --chaos (randomness), --sref (style reference), --iw (image weight), --hd (high-definition mode) |
| Commercial Authorization | General commercial use is supported within the subscription plan (specific terms are subject to the official Terms of Service) |
| Team size | About 60 people, self-profit, no external financing |
| Infrastructure | Based on Google Cloud TPU training and inference |
The decision-making value of the parameter system: Midjourney's competitiveness is not just "being able to produce pictures", but also encapsulating style, composition, randomness, and reference strength into a set of reusable control systems. For the design team, this means that --sref can be used as a "brand style anchor", --stylize controls the depth of style intervention, and --iw determines the influence weight of the reference image. The ability to combine these parameters determines whether the team can turn an accidental good result into a repeatable standardized output.
Midjourney’s users and market recognition
Midjourney’s market recognition is rarely expressed through public corporate client lists or financing news, but through its creator ecosystem, industry savvy and public influence. It has long maintained a high presence in illustration, advertising visuals, conceptual design, independent publishing and social media visual circles, and has been related to the "AI-generated images" events that have entered the public eye many times.
Public Impact Event: In 2022, "Théâtre D'opéra Spatial" generated by Midjourney won the first prize in the Colorado State Fair Digital Art Competition, becoming a landmark event for the first time that AI-generated art won an award in a traditional art competition. In 2023, an AI-generated picture of "Pope Francis wearing a white down jacket" was widely circulated on social media, triggering a global discussion about the authenticity and misleading nature of AI images. In the same year, a fake image of the "Pentagon explosion" generated by Midjourney caused temporary fluctuations in the stock market. Although these events are highly controversial, they also prove from one aspect that the realism of Midjourney's images has reached a level that is difficult to distinguish with the naked eye.
Industry Adoption: The Economist uses Midjourney to create the cover for its June 2022 issue; Italian newspaper Corriere della Sera uses comics created by Midjourney; and several ad agencies and creative studios have incorporated Midjourney into their visual exploration workflows. Architects use it to generate mood boards for early projects, replacing the traditional Google Images search process. Harvard Business Review also published an article exploring how generative AI such as Midjourney can enhance human creativity.
Community Barriers: The official long-term operation of the Discord community, help center and update station (updates.midjourney.com) has formed a very high density of prompts, parameters, cases and secondary creation experience. The unique value of this community is that good workflows are "tried out" by the community, not written in official documents. For newcomers, the parameter combinations, style reference pictures and Prompt templates in the community are more valuable for practical reference than any tutorials.
Publicity Verification: The official claim is to "build the most beautiful AI model in the world." Although this statement is subjective, among the current mainstream Vincentian graphics platforms, Midjourney's output does have a perceptible lead in terms of aesthetic consistency, color tonality, and style cohesion. Its positioning is not "most photo-like", but "best-looking" - this difference determines that its core user group is creative workers who have aesthetic requirements, rather than generalized people who pursue "usable" document illustrations.
Compliance and Litigation Risk: Midjourney faces multiple copyright lawsuits. In January 2023, three artists filed a class-action lawsuit against Stability AI, Midjourney, and DeviantArt, accusing them of using artists' work without authorization to train AI models. In June 2025, Universal Pictures and Walt Disney Company filed a copyright infringement lawsuit against Midjourney, calling it "bottomless plagiarism." In September 2025, Warner Bros. Discovery filed a similar lawsuit, accusing Midjourney of systematically misappropriating its intellectual property, including Superman, Batman and other characters. The direction of these lawsuits will directly affect Midjourney's commercial authorization boundaries and model training strategies, and is a risk dimension that cannot be ignored in procurement evaluations.
Midjourney’s cost advantage: low marginal cost under fixed subscription system
Midjourney adopts a pure subscription system and does not charge based on API Token or the number of images generated. For the usage method of "daily trial and error + style polishing", this billing model naturally lowers the psychological threshold - users do not need to calculate the cost every time they click to generate.
| Plan | Monthly payment price | Annual payment conversion price | Reference number of parallel jobs | Applicable people |
|---|---|---|---|---|
| Basic | $10/month | $96/year (~$8/month) | About 3 concurrencies | Lightweight creation and introductory exploration |
| Standard | $30/month | $288/year (~$24/month) | ~15 concurrencies | High-frequency individual creators |
| Pro | $60/month | $576/year (~$48/month) | ~30 concurrencies | Professional Creators & Studios |
| Mega | $120/month | $1152/year (~$96/month) | ~60 concurrency | Heavy production teams |
C-side/Personal Cost Interpretation: The thresholds for Basic and Standard are moderately low among AI vision tools. The real value is that after subscribing, you can safely draw styles, test parameters, and adjust reference images repeatedly without having to focus on "how much credit is left" every time. For individual creators, the parallel capabilities of the Standard solution are enough to support the daily "multi-directional simultaneous exploration" work mode.
Developer/API cost level: Midjourney is not essentially an API infrastructure product - it does not have public standard API endpoints for developers to integrate secondaryly, nor does it provide a call method that is billed by Token. This means that if a team needs to embed image generation capabilities into its own product pipeline (such as e-commerce batch main image generation), Midjourney is not the most suitable choice. This "no API" positioning is both a boundary and a clear reflection of the cost structure - you don't have to pay for API infrastructure that you don't use.
Enterprise/Studio Cost Interpretation: The actual hidden cost of Pro and Mega is not in the subscription fee itself. The real cost comes from three aspects: Prompt word management and team collaboration - how to maintain a unified Prompt library and style template when multiple people are co-creating; Material screening cost - the greater the generation volume, the longer it takes to manually filter available materials from a large amount of output; Refinement and implementation cost - AI drawings can rarely be delivered directly, and brand finalization usually requires designers to do secondary refinement in Photoshop. These hidden costs are often multiples of the subscription fee itself.
The free truth: Midjourney did not use "free credits" as a means of acquiring customers from the beginning. Its strategy is more like "pay first, experience later". This means that it is not friendly enough for light early adopters, but it is clearer for users who are serious about visual output - you don’t have to guess when the free quota will be used up, everything is transparently visible in the subscription plan.
Midjourney’s main features
Text to Image Generation: Quickly generate multi-style visual drafts through natural language prompt words. V8.1 uses Standard Resolution by default, which is faster than V7’s draft mode and suitable for rapid iterative exploration. HD mode produces approximately 2K resolution images without additional upscaling.
Image Prompts and Image Weight: Users upload reference images and use the --iw parameter to control the degree of influence of the reference images on the final output. The higher the --iw value, the closer the output is to the structure and details of the reference image; the lower the value, the more word-dominated the output will be. V8.1 brings back this feature that was extremely popular in the V7 community.
Style Reference / --sref: Upload a reference image as a style guide, and Midjourney extracts its tone, texture, and overall mood and applies it to the newly generated image. In V8.1, the official clearly stated that "Moodboards and sref are now super stable, and the effect will make you love it". This function is a core tool for brand style unification.
Character Reference (--cref): Upload a character image, and the system will use it as a reference for the character's appearance to maintain character consistency in multiple outputs. This is crucial for projects such as comics, picture books, and series illustrations that require a unified character identity.
Vary and Upscale: Generate variations or upscale resolutions based on existing results, without the need to generate them from scratch each time. V8.1 adds a new "Run as HD" button, which can convert standard resolution jobs to HD mode for regeneration with one click.
Parametric control system: Midjourney's parameter system is its core barrier that distinguishes it from most Vincent diagram tools. The core parameters include: --ar (aspect ratio), --stylize (stylization intensity, range 0-1000), --chaos (result randomness, range 0-100), --v (model version selection), --sref (style reference), --iw (image weight), --hd (high-definition mode). The combined use of these parameters allows the team to solidify an accidental style discovery into a reusable parameter template.
Web Editor: The Web editor launched since V6.1 integrates image editing, panning (Pan), zooming (Zoom Out), regional variation (Vary Region) and local redrawing (Inpainting) into a unified interface. Conversation synchronization between Discord channels and web rooms is also supported.
Expert's perspective - Hidden linkage: The real engineering value of Midjourney does not lie in any isolated function, but in the "prompt word → reference image → parameter → variant → amplification → style solidification" forming a complete exploration package. Specifically: first use text prompt words to divergence direction → use --sref to anchor the style tone → control the degree of style intervention and randomness through --stylize and --chaos → use variants to fine-tune in the confirmed direction → use HD mode to produce the final high-definition version → save the final confirmed parameter combination as a team template. The more mature this relationship is, the faster the team can launch new projects.
Midjourney’s model and version evolution
The version evolution of Midjourney is different from the "parameter scale competition" of basic large models: the value of each upgrade is rarely reflected in the benchmark ranking, but more reflected in the dimensions that creators really care about: "whether the aesthetic is more stable", "whether the structure is more accurate", "whether the reference picture is more obedient", and "whether the style is more reproducible".
Mainline release
-
V8.1 Alpha (2026-04-14): Current default model. The aesthetic style returns to the stability tone of V7; the HD mode speed is increased by 3 times and the cost is reduced by 3 times; the standard resolution speed is increased by 50% and the cost is reduced by 25%. The standard resolution at full quality is faster than the V7 draft mode. Reintroduced the Image Prompts and Image Weights (
--iw) functions. Added Prompt Shortener and updated Describe tool to generate longer detailed prompts. Draft Mode and Random Styles have been launched. -
V8 Alpha (2026-03-17): The new generation model mainline is opened, the model base is strengthened, and the architecture is prepared for subsequent editing, amplification and video link upgrades. V8.0 has been taken offline a few weeks after the release of V8.1.
-
V7 (2025-04-01): Long-term main version. Emphasis on highly consistent style expression and parameter control experience.
--srefand Moodboards reach a highly available state in this release and become core workflows widely used by the community. -
V6.1 (2024-07-30): Further optimize the realism quality and prompt word following ability. At the same time, the web editor was officially launched, integrating image editing, panning, zooming, regional variation and local redrawing into a single web interface, marking Midjourney's key transformation from "Discord subsidiary tool" to "independent web platform".
Candidate Verification
-
V6 (2023-12-21): A model trained from scratch, which took 9 months. Significantly improve text rendering capabilities (the ability to generate readable text in images) and prompt word text compliance capabilities. This is the starting point for the Web Alpha version.
-
V5.2 (2023-06): Introducing the "Aesthetic System" and "Zoom Out" image expansion function to support the outward generation of surrounding context based on existing images. The Vary (Region) region variant feature is released in this version.
-
V5.1 (2023-05): More stylized than V5, while introducing a more "literal" RAW mode. During the same period, it switched from a keyword blocking system to an AI-driven moderation system that allows for context-sensitive word usage.
-
V5 (2023-03-15): A clear breakthrough has been made in the realism and "uncanny valley" issues, and the image quality has improved dramatically.
How to read the version: The main version of Midjourney is usually first released in alpha form on alpha.midjourney.com, and is fully pushed to the main site and Discord after a few weeks of community testing. The best sources of information for following releases are updates.midjourney.com and the official Discord #announcements channel.
Niji Branch: Midjourney also has a Niji Journey branch specifically for anime style. Niji V6, released in January 2024, and Niji V7, released in January 2026, improve anime continuity, prompt word understanding, text rendering, and --sref performance.
Midjourney’s technical advantages
Mechanism → Effect → Scene: Midjourney has long focused model tuning on visual aesthetics and style cohesion, rather than solely pursuing the "more realistic" dimension. Its training strategy and loss function design tend to favor outputs that perform better in color harmony, composition balance, and light and shadow coordination. The effect is that it is easier to produce films in scenes that require "unification of temperament" such as concept posters, brand mood maps, illustrations and fashion visuals. Applicable scenario is that Midjourney has more obvious advantages for any task that requires "delivering a feeling" rather than "accurately restoring an object".
Design philosophy of parameter system: --ar, --stylize, --chaos, --sref, --iw and other parameters are not simple numerical sliders, but structured control of the model generation space. --stylize controls the "degree of involvement of the model in aesthetics" - the lower the value, the more strictly the model follows the original text of your prompt word; the higher the value, the more the model tends to "beautify" the results with its own aesthetic preferences. --chaos controls the "output diversity" - low chaos means the output is closer each time, and high chaos means the difference is greater. This system allows the team to gradually consolidate contingency into reusable methodologies. For the studio, this is more important than simply producing a few more good-looking pictures, because reusability is the real productivity.
Google Cloud TPU Infrastructure: Starting from version V4, Midjourney is based on Google Cloud TPU for model training and inference. This option reduces infrastructure management costs compared to building your own GPU cluster, while leveraging Google's global network to accelerate the response time of the image generation service.
Evolution of Content Moderation Mechanism: In the early days, Midjourney relied on a keyword blocking system (banning certain names, religious words, etc.), which caused a lot of controversy about censorship. In May 2023, Midjourney switched to an AI-driven moderation system that can analyze the context of prompt words to decide whether to block, rather than simply blocking the word. For example, users can now generate portraits containing "Xi Jinping", but the system will prevent controversial scenes such as the character's arrest from being generated. This mechanism finds an engineered balance between preserving creative freedom and complying with red lines.
Boundary - What it is not suitable for: What Midjourney is least good at is enterprise scenarios with strong processization, strong material asset management, and strong localized compliance. It is also not suitable for use as a structured image editing system. It is not suitable for producing accurate product renderings (which require full control of size and material), is not suitable for scenarios such as medicine/engineering that require precise structural expression, and is not suitable for organizations that have strict requirements for data sovereignty. It is positioned as a "visual front-end engine", not a complete design collaboration platform.
How to use Midjourney
Midjourney provides two core entrances, corresponding to different workflow preferences.
Web (alpha.midjourney.com): This is the currently recommended main entrance. Supports visual management of generated records, creation of Moodboards, management and reuse of style references, and batch operations. All functions of V8.1 (HD mode Image Prompts, Prompt Shortener, Describe tool) can be used directly on the Web. The web interface also supports one-click switching of "Run as HD", as well as functions such as Draft Mode and Random Style.
Discord Bot: Use the /imagine command in the official Discord server, or invite Midjourney Bot to your own server. Ideal for community collaboration, instant feedback, and rapid testing. The core advantage of the Discord side is the community ecology - you can see other people's Prompt parameters and results in real time in channels such as #tutorial and #show-and-tell.
Suggested Workflow:
- Use simple prompt words (best in English) on the web to conduct the first round of direction exploration.
- After locking the direction, anchor the style tone through
--srefor Moodboard - Use
--stylizeand--chaosto fine-tune the degree of style involvement and diversity - Use Vary to refine in the confirmed direction
- Use HD mode or "Run as HD" button to generate final HD version
- Save the final combination of prompt words and parameters as a template that can be reused by the team
3 minutes to get started quickly: Visit alpha.midjourney.com → Enter a simple English prompt word (such as "minimalist coffee shop interior, warm lighting") → Observe the 4 outputs → Click the Vary button on one of them to generate a variant → Try to add --ar 16:9 to change the frame → Add --sref and add a reference picture. This process is the fastest way to determine whether you like its aesthetics and control system.
Chinese user usage boundaries: Midjourney provides services to users around the world, but the main interface, documentation and community communication are still mainly in English. The best practice for Chinese users is to use English prompt words, or mixed Chinese and English prompt words (English mainly, Chinese as a supplementary description style). Payment availability, access experience and account availability are subject to evaluation based on local network and billing conditions. There is currently no official independent announcement of localized products and localized billing systems in mainland China.
Midjourney product pricing
Midjourney is purely subscription-based and does not have a pay-as-you-go or free model. There are four core tiers, and annual payments usually offer a discount of about 20%.
| Plan | Monthly payment | Annual payment (average monthly) | Parallel upper limit reference | Core differentiation capabilities |
|---|---|---|---|---|
| Basic | $10/month | ~$8/month | About 3 concurrencies | Limited quick mode duration, suitable for entry-level exploration |
| Standard | $30/month | ~$24/month | About 15 concurrencies | Unlimited slow mode + appropriate fast mode, preferred by high-frequency individual creators |
| Pro | $60/month | ~$48/month | ~30 concurrencies | More Quick Mode duration, Stealth Mode generation |
| Mega | $120/month | ~$96/month | ~60 concurrency | Maximum fast mode duration, suitable for heavy production teams |
Billing Instructions: Detailed specifications such as fast mode duration limit, number of concurrencies, and stealth mode availability for each plan are subject to real-time information on the official website subscription page. Annual payment plans are usually more cost-effective than monthly payments. The scope of commercial use rights for images generated within all subscription plans must be implemented in accordance with the official Terms of Service, and may differ between packages.
Refunds and Terms: Please refer to the official billing page and Terms of Service. It is recommended to read the terms of commercial use carefully before purchasing, especially the restrictions on use during the period affected by copyright litigation.
Application scenarios of Midjourney
Advertising and Brand Creativity: Quickly generate multiple versions of visual directions during the proposal stage to help customers confirm the style and tone before committing to formal design. In the traditional process, a visual proposal for one direction may require a designer to produce a concept draft in 1-2 days; Midjourney can compress "getting a discussable direction" into 10-30 minutes. Key points of verification: Confirm whether the final selected style can be solidified through --sref and Moodboard to ensure the consistency of subsequent batch output.
Illustration and concept design: used for pre-production of character design, scene atmosphere map, and world view exploration. The character reference (--cref) function can maintain the consistency of the appearance of the same character in multiple pictures, which is especially important for comics, picture books, and game art. Verification focus: Test the character restoration degree of --cref in different scenes and different angles.
E-commerce and social media content: Low-cost batch production of visual materials with a unified style. For example, cross-border e-commerce product scene pictures, social media festival marketing posters, and grass-growing content header pictures. Key points of verification: Evaluate the efficiency of material screening - the greater the volume of production, the easier it is for the labor costs of screening and refinement to exceed the subscription fee itself.
Film and television storyboards and proposals: Quickly transform creative text into visual storyboard images or concept proposal drawings. Midjourney's strong stylization ability is an advantage here - it does not provide "accurate storyboards", but "visualization of atmosphere and emotions", helping the team align their visual imagination before formal production. Verification Points: Midjourney is not suitable for scenes that require precise control of lens parameters (focal length, aperture, camera height).
Quantitative deduction of cost reduction and efficiency improvement (Rule D mandatory):
- Writing weekly reports and matching pictures for new media operations: In the traditional process, operators need to search PS from the gallery to modify pictures, or ask designers to produce pictures. A single picture takes about 30-60 minutes. Midjourney can compress the time "from copywriting to usable images" to 5-10 minutes, improving efficiency by about 5-10 times.
- Brand designer produces concept drawings: In the traditional process, it takes 0.5-2 days from receiving the brief to producing concept drawings in 3-5 directions. Using Midjourney can compress the "first round direction output" to 15-45 minutes, improving efficiency by about 8-15 times.
- Pre-game art concept exploration: In the traditional process, character/scene concept exploration requires artists to hand-draw multiple versions of sketches, each version taking 2-6 hours. Midjourney can generate 20-40 variants for screening in 10-20 minutes, which is about 10-20 times more efficient.
(The above deduction is based on industry experience and is an unofficial commitment. The actual efficiency improvement depends on the team's proficiency in the parameter system and screening efficiency.)
Dimensionality reduction attack scenarios: style mood maps, concept posters, character worldview exploration, direction maps before customer proposals, social media batch visual content - these are the most comfortable use areas of Midjourney.
Not applicable to scenarios: Scenarios that require accurate product rendering (the position of each screw is controllable), medical/engineering structural drawings, enterprise-level material asset management, and strong compliance local privatization deployment are not suitable for using Midjourney as the core production system.
Midjourney is suitable for people
Individual Creator: Illustrator, graphic designer, content creator, independent publisher. Suitable for individuals who require high-frequency visual output and have requirements for aesthetic quality. The prerequisite is to have basic English prompt word writing ability and willingness to adjust parameters.
Small creative team: advertising agency, brand studio, game art team, short video MCN. The value of Midjourney lies in increasing the team’s visual exploration efficiency by an order of magnitude. The cost required is not the subscription fee, but the team's time to accumulate the Prompt library and style templates.
Education and Research Field: Teaching tool for visual communication major, experimental platform for AI and creativity research, and practical tool for creative workshops. Ideal for teaching and research exploring AI-assisted creation.
Advertising and marketing practitioners: social media operations, e-commerce designers, brand planning. Midjourney has basically no rivals in the "quickly produce a lot of good-looking but not necessarily accurate visual content" scenario.
Dissuaded/not applicable to people:
- Industrial designers and product engineers who need accurate product renderings and engineering structural drawings
- Organizations that require strict privatization deployment or data sovereignty compliance (such as financial institutions, military industrial enterprises)
- Development teams who need to embed image generation into automated business pipelines (no public API)
- Corporate legal departments that have zero tolerance for AI copyright risks (need to pay close attention to the progress of multiple class actions)
- Content creators who have no English foundation and are unwilling to learn the parameter system
Summary and Outlook of Midjourney
The long-term value of Midjourney lies in the combination of "aesthetic expression + parameter control + community experience", not just the quality of a single rendering. It turns AI visual tools into a way of working that is close to creative habits - designers are not dealing with a "black box of random drawings", but are collaborating with a set of controllable parameter languages and models. This is also the fundamental reason why it still maintains strong brand power and user loyalty after going through multiple rounds of iterations from V5 to V8.
As the V8 series continues to advance, there are several directions worth paying attention to: the stability and functional perfection of the official version of V8 (especially the upgrade of post-processing capabilities such as editing Inpainting and Outpainting); whether the web workflow will replace Discord as the main entrance; whether Midjourney's exploration in the video generation and hardware direction (officially announced hardware projects and Medical direction) can form a second growth curve.
Unsuitable boundary reiterated: If what the team needs most is high-quality creative divergence and style exploration, Midjourney is currently one of the most mature options. But if the organization needs strong governance, strong API, and strong enterprise integration, it is more suitable as a "pre-creation engine" rather than the only production platform. For teams that are already using Midjourney, it is recommended to establish supporting prompt word management specifications and production quality standards to control hidden costs within an acceptable range.
Procurement/Adoption Risk Assessment:
- Uncertainty about copyright litigation: Large copyright parties such as Disney, Youball Pictures, and Warner Bros. Discovery have filed lawsuits against Midjourney. If a court rules that Midjourney infringes copyright, it could affect the legality of its training data, the usability of its models, and the scope of commercial uses of its user-generated content. Businesses should ask their legal department to assess the potential impact of these lawsuits before making a purchase.
- No Private Deployment Option: Midjourney is currently only available as a SaaS subscription and cannot be deployed on-premises. Organizations that have rigid requirements for data sovereignty are not suitable to use Midjourney as their core production tool.
- No public API: Cannot be integrated into own product lines via API like DALL-E or Stable Diffusion. If you need to automate your image generation pipeline, consider other API-capable alternatives.
- Recommended pilot method: First try the Standard solution in a small team (2-5 people) for 1-2 months to verify its efficiency improvement and hidden costs in real workflows, and then decide whether to expand to the entire team.
Related tools: Midjourney, stable-diffusion
Comparison of competing products
| Comparison dimensions | The tool | Competitor A | Competitor B |
|---|---|---|---|
| Core Differences | — | — | — |
| Price | — | — | — |
| Target User | — | — | -- |
Version Info
- V8.1 Alpha :V8.1 Alpha provides a more stable aesthetic style, speeds up and reduces costs in HD mode, and restores image prompts and a stronger Prompt tool chain.
- V8 Alpha :V8 trunk preview version, strengthens the model base and is suitable for subsequent editing, amplification and video link upgrades.
- Version 7 :As the long-term main version, V7 emphasizes high consistency style and parameter control experience.
- Version 6.1 :V6.1 further optimizes the realism quality and prompt word following ability.
User Reviews