• About
  • Advertise
  • Privacy & Policy
  • Contact
HK Businesswire
  • Home
  • News
    • All
    • Business
    • Politics
    • PR Newswire
    • Science
    • World

    Hong Kong’s 5-year plan sees over 2,500 submissions

    Ride-hailing permit grants ‘to be fair, transparent’

    Agoda Marks 11th Tech Camp Day with Agentic AI Focus

    Agoda Marks 11th Tech Camp Day with Agentic AI Focus

    Global Times: Facing the questions of the AI era, the world looks to China for solutions

    Global Times: China sends fresh signal on global AI cooperation at WAIC

    Global Times: Head-of-state diplomacy shines at WAIC, fostering ties and advancing global governance consensus

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • PR Newswire
  • Business
  • World
  • Entertainment
  • Sports
  • Tech
    • All
    • Apps
    • Gadget
    • Mobile
    • Startup

    Alipay Launches AI-Powered Version ‘Abao’ to Streamline Services

    Xiaohongshu Prepares Confidential Hong Kong IPO Filing

    SpaceX Raises $75 Billion in Historic IPO Amid $350 Billion Investor Demand

    Chinese firms double down on tech: Xiaomi, Haier

    Xiaomi Launches MiMo Code AI Programming Assistant to Enter Coding Agent Market

    Apple Unveils Overhauled Siri AI and Major OS Updates at WWDC 2026

    OpenAI launches AI browser Atlas

    OpenAI Files Confidentially for IPO Amid Intensifying AI Competition

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
  • Feature
No Result
View All Result
  • Home
  • News
    • All
    • Business
    • Politics
    • PR Newswire
    • Science
    • World

    Hong Kong’s 5-year plan sees over 2,500 submissions

    Ride-hailing permit grants ‘to be fair, transparent’

    Agoda Marks 11th Tech Camp Day with Agentic AI Focus

    Agoda Marks 11th Tech Camp Day with Agentic AI Focus

    Global Times: Facing the questions of the AI era, the world looks to China for solutions

    Global Times: China sends fresh signal on global AI cooperation at WAIC

    Global Times: Head-of-state diplomacy shines at WAIC, fostering ties and advancing global governance consensus

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • PR Newswire
  • Business
  • World
  • Entertainment
  • Sports
  • Tech
    • All
    • Apps
    • Gadget
    • Mobile
    • Startup

    Alipay Launches AI-Powered Version ‘Abao’ to Streamline Services

    Xiaohongshu Prepares Confidential Hong Kong IPO Filing

    SpaceX Raises $75 Billion in Historic IPO Amid $350 Billion Investor Demand

    Chinese firms double down on tech: Xiaomi, Haier

    Xiaomi Launches MiMo Code AI Programming Assistant to Enter Coding Agent Market

    Apple Unveils Overhauled Siri AI and Major OS Updates at WWDC 2026

    OpenAI launches AI browser Atlas

    OpenAI Files Confidentially for IPO Amid Intensifying AI Competition

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
  • Feature
No Result
View All Result
HK Businesswire
No Result
View All Result
Home News PR Newswire

Vidu Q1 Model Update Unveils Multi-Reference Feature, Supporting Up to Seven Image Inputs

PR Newswire by PR Newswire
8 July 2025
in PR Newswire
0
Vidu Q1 Model Update Unveils Multi-Reference Feature, Supporting Up to Seven Image Inputs
0
SHARES
5
VIEWS
Share on FacebookShare on Twitter

SINGAPORE, July 8, 2025 /PRNewswire/ — Vidu, the flagship product of ShengShu Technology and a pioneer in generative AI video, today announced a major update to its latest Vidu Q1 model, offering a new advanced ‘Reference-to-Video’ feature, powered by semantic understanding, which in an industry-first can generate videos at scale from up to seven image inputs.


Producing complex, multi-character films with AI is quickly becoming a reality. But until now, maintaining visual consistency across multiple scenes and videos has been one of the field’s most difficult challenges. Characters would shift subtly between shots, objects would change, and continuity would break down entirely from one video to the next.

Vidu Q1 changes that with its Reference-to-Video feature. This breakthrough allows creators to preserve a high level of consistency from the first video to the videos generated thereafter, from character appearance and behavior to background elements and props. The model understands and tracks visual identity across frames, so new additions to a scene won’t disrupt the continuity of what’s already been established.

This profound capability directly underpins how this update to Vidu’s dynamic Q1 generative video model marks a transformative milestone. Scenes that once required tens of millions of dollars and months of production can now be cut down to hundreds of dollars and a matter of minutes. In fact, generating a 5-second 1080p video clip with Vidu Q1’s Reference-to-Video feature can cost as little as $0.14. This is less than the cost of a can of soda for a high-definition video.

Imagine recreating the iconic Battle of Helm’s Deep from The Lord of the Rings, which originally took 120 days and hundreds of actors and extras. Vidu Q1 brings us closer to achieving a similar scale on a shoestring production budget, and all within the comfort of your home. It signals a paradigm shift: AI models are becoming sophisticated and perceptive enough to synthesize cinematic-scale detail with minimal human input and visual cues, redefining the boundaries of what’s possible in filmmaking.

Inferring the Unseen: Generating Objects Without Reference Images

At the heart of the Reference-to-Video feature is Vidu Q1’s enhanced semantic understanding engine. By understanding how reference images relate to text prompts, Vidu Q1 can automatically infer missing visual elements and generate key objects described in the prompt, even if they aren’t explicitly present in the input images.

For example, a creator might upload an image of a man, a bird, and a cityscape, then prompt: “The man plays a violin while the bird lands on his shoulder in the city at sunset.” Even if no violin image is provided, Vidu’s semantic core generates and integrates a violin seamlessly, preserving visual consistency and narrative clarity throughout the clip.

With this breakthrough, creators no longer face steep technical hurdles when attempting to create complex scenes. A simple text prompt is interpreted and realized by the model’s robust semantic layer, removing the need to upload every individual prop or element. This way, users can focus on storytelling and visual impact, letting Vidu handle asset generation and making sure each scene is coherent.

Unlocking Next-Level Consistency Multi-Reference for Up to Seven Images

Vidu Q1’s expanded multi-image reference capability, which supports up to seven reference images per video sequence, is a major leap forward for AI filmmaking. Creators can build visually richer scenes featuring multiple characters, props, or backgrounds, without ever needing them in the same room. This is more than just a feature. It’s a creative unlock that pushes generative video closer to the complexity of traditional film production using only prompts and reference visuals.

“This update breaks through the limits of what creators thought they could do with AI video. We’re getting closer to enabling users to create fully realized scenes, complete with a detailed cast of characters, objects, and backgrounds, by expanding multi-image referencing to support up to seven inputs,” said Luo Yihang, CEO at ShengShu Technology. “Combined with advanced semantic understanding, this marks a major step toward bridging pure imagination with precise execution. This will allow users to craft scenes with deliberate structure and consistency, progressing from isolated clips to fully formed, narratively cohesive videos.”

In addition, the references can be saved to a personal library of images from which they can be repurposed for future projects for generating more detailed scenarios with layered visual elements, while maintaining stable continuity frame to frame. While the Vidu Q1 model supports up to seven images for now, this is just the start. This capability is continuously being optimized to achieve greater stability.

Learn more about Vidu Q1 here: https://vidu.com/

About ShengShu Technology

Founded in March 2023, ShengShu Technology is a world-leading artificial intelligence company, specializing in the development of Multimodal Large Language Models. Driven by innovation, the company delivers cutting-edge MaaS and SaaS products that revolutionize creative production by enabling smarter, faster, and more scalable content creation. With its flagship video generation platform Vidu, ShengShu Technology’s solutions have reached more than 200 countries and regions around the world, spanning fields including interactive entertainment, advertising, film, animation, cultural tourism, and more.

Tags: prnewswire
PR Newswire

PR Newswire

PR Newswire is the industry’s leading press release distribution partner with an unparalleled global reach of more than 440,000 newsrooms, websites, direct feeds, journalists and influencers and is available in more than 170 countries and 40 languages. From our award-winning Content Services offerings, integrated media newsroom and microsite products, Investor Relations suite of services, paid placement and social sharing tools, PR Newswire has a comprehensive catalog of solutions to solve the modern-day challenges PR and communications teams face. For 70 years, PR Newswire has been the preferred destination for brands to share their most important news stories across the world.

Read More

Singtel Receives Four Frost & Sullivan 2026 Recognitions for Leadership in Enterprise Connectivity, Cybersecurity, and Digital Transformation

19 July 2026
Emdoor Launches “Ailyn” AI Hub at WAIC 2026: Unifying Intelligence Across Every Device

Emdoor Launches “Ailyn” AI Hub at WAIC 2026: Unifying Intelligence Across Every Device

19 July 2026
  • Trending
  • Comments
  • Latest
CyberLogitec Wins Smart Terminal Project at TTIA in Europe, Delivering TOS and Digital Twin Solutions

CyberLogitec Wins Smart Terminal Project at TTIA in Europe, Delivering TOS and Digital Twin Solutions

14 July 2026
Subsidised housing ballots drawn

Subsidised housing ballots drawn

10 July 2026
Xiao Noodles Posts Maiden Annual Results: Revenue and Net Profit Jump in 2025 as ESG Efforts Drive Long-Term Value

Xiao Noodles Posts Maiden Annual Results: Revenue and Net Profit Jump in 2025 as ESG Efforts Drive Long-Term Value

29 April 2026

HBO Max Restores Traditional Chinese Subtitles in Hong Kong After User Backlash

31 January 2026

Hong Kong’s 5-year plan sees over 2,500 submissions

18 July 2026

Ride-hailing permit grants ‘to be fair, transparent’

18 July 2026
Agoda Marks 11th Tech Camp Day with Agentic AI Focus

Agoda Marks 11th Tech Camp Day with Agentic AI Focus

18 July 2026

Global Times: Facing the questions of the AI era, the world looks to China for solutions

18 July 2026

Recent News

Hong Kong’s 5-year plan sees over 2,500 submissions

18 July 2026

Ride-hailing permit grants ‘to be fair, transparent’

18 July 2026
Agoda Marks 11th Tech Camp Day with Agentic AI Focus

Agoda Marks 11th Tech Camp Day with Agentic AI Focus

18 July 2026

Global Times: Facing the questions of the AI era, the world looks to China for solutions

18 July 2026
HK Businesswire

Stay ahead with the latest insights on Hong Kong’s economy, finance, and investments. From market trends to policy updates, we bring you in-depth analysis and expert opinions.

📩 Subscribe to our newsletter for exclusive updates.
📍 Follow us on social media for real-time news.
📧 Contact us: info@hongkong-invest.com

Follow Us

  • About
  • Advertise
  • Privacy & Policy
  • Contact

© 2025 by HKBusinesswire.com

No Result
View All Result

© 2025 by HKBusinesswire.com