跳至内容
  • 首页
  • 资讯
  • 资源下载
  • 行业方案
  • Job招聘
  • Paper论文
  • Patent专利
  • 映维会员
  • 导航收录
  • 合作
  • 关于
  • 微信群
  • All
  • XR
  • CV
  • CG
  • HCI
  • Video
  • Optics
  • Perception
  • Reconstruction

FlexGen: Flexible Multi-View Generation from Text and Image Inputs

编辑:广东客   |   分类:CV   |   2025年3月27日

Note: We don't have the ability to review paper

PubDate: Otc 2024

Teams:HKUST(GZ)1 HKUST2 Quwan3

Writers:Xinli Xu, Wenhang Ge, Jiantao Lin, Jiawei Feng, Lie Xu, HanFeng Zhao, Shunsi Zhang, Ying-Cong Chen

PDF:FlexGen: Flexible Multi-View Generation from Text and Image Inputs

Abstract

In this work, we introduce FlexGen, a flexible framework designed to generate controllable and consistent multi-view images, conditioned on a single-view image, or a text prompt, or both. FlexGen tackles the challenges of controllable multi-view synthesis through additional conditioning on 3D-aware text annotations. We utilize the strong reasoning capabilities of GPT-4V to generate 3D-aware text annotations. By analyzing four orthogonal views of an object arranged as tiled multi-view images, GPT-4V can produce text annotations that include 3D-aware information with spatial relationship. By integrating the control signal with proposed adaptive dual-control module, our model can generate multi-view images that correspond to the specified text. FlexGen supports multiple controllable capabilities, allowing users to modify text prompts to generate reasonable and corresponding unseen parts. Additionally, users can influence attributes such as appearance and material properties, including metallic and roughness. Extensive experiments demonstrate that our approach offers enhanced multiple controllability, marking a significant advancement over existing multi-view diffusion models. This work has substantial implications for fields requiring rapid and flexible 3D content creation, including game development, animation, and virtual reality. Project page: this https URL.

本文链接:https://paper.nweon.com/16252

您可能还喜欢...

  • Semantic Scene Completion via Integrating Instances and Scene in-the-Loop

    2021年07月02日 映维

  • 254de596fcc983d91bea9dfbdb27adc4-thumb-medium

    The Design of Real-Time Digital Clothing Projection System

    2020年11月06日 映维

  • 20d5d946fcaf624ee39cd4a124ddb1d0-thumb-medium

    Vision-based Pose Estimation for Augmented Reality : A Comparison Study

    2020年07月31日 映维

关注:

最新AR/VR行业分享

  • ★ 映维日报:苹果或为M5版Vision Pro推出新头带,高通和中移动发布弱视儿童PICO VR关爱计划 2025年10月21日
  • ★ 高通和中移动发布“睛彩无界”弱视儿童VR关爱计划,现场演示基于PICO的解决方案 2025年10月21日
  • ★ 苹果推送visionOS 26.1开发者预览版23N5042a更新 2025年10月21日
  • ★ 广州执信中学计划以230万元采购VR/STEAM课室设备 2025年10月21日
  • ★ 东台市第一小学计划以160万建设VR实验室、数字探究实验室 2025年10月21日

最新AR/VR专利

  • ★ Sony Patent | System and methods for electronic gaming control and performance normalization using artificial intelligence 2025年10月16日
  • ★ Microsoft Patent | Utilizing noise threshold conditions for controlling scanning mirrors 2025年10月16日
  • ★ Samsung Patent | Wearable device for controlling at least one virtual object according to attributes of at least one virtual object, and method for controlling same 2025年10月16日
  • ★ Meta Patent | Techniques for interactive visualization for workspace awareness in collaborative authoring of metaverse environments, and systems and methods of use thereof 2025年10月16日
  • ★ Apple Patent | Systems, methods, and graphical user interfaces for modeling, measuring, and drawing using augmented reality 2025年10月16日

最新AR/VR行业招聘

  • ★ Microsoft AR/VR Job | High Performance Compute, Director 2025年6月5日
  • ★ Microsoft AR/VR Job | Data Center Technician/ Technicien de Centre de Données 2025年6月3日
  • ★ Microsoft AR/VR Job | Senior Product Designer 2025年5月16日
  • ★ Apple AR/VR Job | AirPlay Audio Engineer 2025年3月27日
  • ★ Apple AR/VR Job | iOS Perception Engineer 2025年3月27日
  • 首页
  • 资讯
  • 资源下载
  • 行业方案
  • Job招聘
  • Paper论文
  • Patent专利
  • 映维会员
  • 导航收录
  • 合作
  • 关于
  • 微信群

联系微信:ovalics

版权所有:广州映维网络有限公司 © 2025

备案许可:粤ICP备17113731号-2

备案粤公网安备:44011302004835号

友情链接: AR/VR行业导航

读者QQ群:251118691

Quest QQ群:526200310

开发者QQ群:688769630

Paper