Skip to content
Conference

From text to immersive 3D: an end-to-end pipeline for automated retrieval-augmented virtual room generation

Aug 2026 · International Conference on Computer Graphics and Virtuality · Vol 14315, pp. 143150D - 143150D-12 · 0 citations · 20 references
Engineering

TL;DR

Results demonstrate that immersive WebXR representations substantially strengthen user engagement and understanding, offering a scalable pathway for employer branding and digital recruitment, and point to interactivity as a key avenue for future work.

Abstract

This paper introduces a retrieval-augmented Text-to-WebXR pipeline that converts free text into interactive, browser-based 3D environments. Leveraging embedding-based retrieval, cross-encoder reranking, and a structured scene schema with constraint-aware placement, the approach generated one validated room description for each of the 229 firms. Expert evaluation indicates consistently high quality (M=4.46/5), with near-perfect scores in consistency (4.99) and completeness (4.88). A user study with 204 participants shows that 79.6% perceive 3D environments as a meaningful enhancement to textual profiles; participants with less prior 3D experience found the representations more helpful, indicating a novelty effect. These results demonstrate that immersive WebXR representations substantially strengthen user engagement and understanding, offering a scalable pathway for employer branding and digital recruitment, and point to interactivity as a key avenue for future work.

View source

Similar papers

Review Open access Aug 2026

A Review of Retrieval-Augmented Generation Technology

Retrieval-augmented generation has emerged as a core technological paradigm for addressing the bottlenecks of hallucinations and knowledge lag in large language models. However, many existing reviews focus on a single technical branch or vertical application scenario, making only scattered references to hardware, evalu...

Peng Jiang, Xiao-Sheng Cai · 0 citations
Review Open access Sep 2026

Design recommendations and future prospects for text legibility in augmented reality

This paper provides a systematic review of the literature on text legibility in augmented reality (AR) using the PRISMA methodology, with an emphasis on head-mounted displays (HMD) and screen devices. Based on the literature review, we propose design recommendations for legibility in AR. The analysis was carried out us...

A. Jović, Ivana Žiljak Stanimirović, M. Mikota et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Text-Driven Artistic Staging: 3D Posing, Lighting, and Camera References from Paintings

This work constructs 11,911 text--staging pairs from 2,328 figurative paintings by reconstructing SMPL bodies, estimating low-frequency illumination, recovering camera parameters, and pairing each scene with ArtEmis descriptions, demonstrating the feasibility of generating editable, emotionally conditioned 3D staging r...

Yun-Ge Wen · 0 citations
Preprint Aug 2026

DesignAgent3D: Interactive 3D Scene Editing via Designer-like Multimodal Reasoning

DesignAgent3D is presented, an interactive multimodal agentic framework that reformulates 3D scene editing as a designer-like Plan-Perceive-Act paradigm, delivering superior semantic intent alignment, impeccable spatial localization accuracy, and high-fidelity multi-view consistency.

Xiu-Jin Liu, Tian-Yu Yang, Yilun Zhao et al. · 0 citations
Open access Aug 2026

Text-to-hierarchical three-dimensional scene generation: a new approach for layered three-dimensional modeling from natural language

An end-to-end hierarchical framework for text-to-3D scene generation that synergistically integrates state-of-the-art components for video synthesis and mesh reconstruction is introduced, offering a powerful solution for applications, such as virtual reality and digital twins.

Zuan Gu, Tian-Han Gao, Lang-Xu Zhao et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.