Graphics

2025-04-07 | | Total: 6

#1 NeRFlex: Resource-aware Real-time High-quality Rendering of Complex Scenes on Mobile Devices [PDF] [Copy] [Kimi] [REL]

Authors: Zhe Wang, Yifei Zhu

Neural Radiance Fields (NeRF) is a cutting-edge neural network-based technique for novel view synthesis in 3D reconstruction. However, its significant computational demands pose challenges for deployment on mobile devices. While mesh-based NeRF solutions have shown potential in achieving real-time rendering on mobile platforms, they often fail to deliver high-quality reconstructions when rendering practical complex scenes. Additionally, the non-negligible memory overhead caused by pre-computed intermediate results complicates their practical application. To overcome these challenges, we present NeRFlex, a resource-aware, high-resolution, real-time rendering framework for complex scenes on mobile devices. NeRFlex integrates mobile NeRF rendering with multi-NeRF representations that decompose a scene into multiple sub-scenes, each represented by an individual NeRF network. Crucially, NeRFlex considers both memory and computation constraints as first-class citizens and redesigns the reconstruction process accordingly. NeRFlex first designs a detail-oriented segmentation module to identify sub-scenes with high-frequency details. For each NeRF network, a lightweight profiler, built on domain knowledge, is used to accurately map configurations to visual quality and memory usage. Based on these insights and the resource constraints on mobile devices, NeRFlex presents a dynamic programming algorithm to efficiently determine configurations for all NeRF representations, despite the NP-hardness of the original decision problem. Extensive experiments on real-world datasets and mobile devices demonstrate that NeRFlex achieves real-time, high-quality rendering on commercial mobile devices.

Subjects: Graphics , Computer Vision and Pattern Recognition , Machine Learning , Multimedia , Performance

Publish: 2025-04-04 12:53:33 UTC


#2 Learning Human Perspective in Line Drawings from Single Sketches [PDF] [Copy] [Kimi] [REL]

Authors: Jinfan Yang, Leo Foord-Kelcey, Suzuran Takikawa, Nicholas Vining, Niloy Mitra, Alla Sheffer

Artist-drawn sketches only loosely conform to analytical models of perspective projection. This deviation of human-drawn perspective from analytical perspective models is persistent and well known, but has yet to be algorithmically replicated or even well understood. Capturing human perspective can benefit many computer graphics applications, including sketch-based modeling and non-photorealistic rendering. We propose the first dedicated method for learning and replicating human perspective. A core challenge in learning this perspective is the lack of suitable large-scale data, as well as the heterogeneity of human drawing choices. We overcome the data paucity by learning, in a one-shot setup, from a single artist sketch of a given 3D shape and a best matching analytical camera view of the same shape. We match the contours of the depicted shape in this view to corresponding artist strokes. We then learn a spatially continuous local perspective deviation function that modifies the camera perspective projecting the contours to their corresponding strokes while retaining key geometric properties that artists strive to preserve when depicting 3D content. We leverage the observation that artists employ similar perspectives when depicting shapes from slightly different view angles to algorithmically augment our training data. First, we use the perspective function learned from the single example to generate more human-like contour renders from nearby views; then, we pair these renders with the analytical camera contours from these views and use these pairs as additional training data. The resulting learned perspective functions are well aligned with the training sketch perspectives and are consistent across views. We compare our results to potential alternatives, demonstrating the superiority of the proposed approach, and showcasing applications that benefit from learned human perspective.

Subject: Graphics

Publish: 2025-04-04 00:57:48 UTC


#3 Robust AI-Synthesized Image Detection via Multi-feature Frequency-aware Learning [PDF] [Copy] [Kimi] [REL]

Authors: Hongfei Cai, Chi Liu, Sheng Shen, Youyang Qu, Peng Gui

The rapid progression of generative AI (GenAI) technologies has heightened concerns regarding the misuse of AI-generated imagery. To address this issue, robust detection methods have emerged as particularly compelling, especially in challenging conditions where the targeted GenAI models are out-of-distribution or the generated images have been subjected to perturbations during transmission. This paper introduces a multi-feature fusion framework designed to enhance spatial forensic feature representations with incorporating three complementary components, namely noise correlation analysis, image gradient information, and pretrained vision encoder knowledge, using a cross-source attention mechanism. Furthermore, to identify spectral abnormality in synthetic images, we propose a frequency-aware architecture that employs the Frequency-Adaptive Dilated Convolution, enabling the joint modeling of spatial and spectral features while maintaining low computational complexity. Our framework exhibits exceptional generalization performance across fourteen diverse GenAI systems, including text-to-image diffusion models, autoregressive approaches, and post-processed deepfake methods. Notably, it achieves significantly higher mean accuracy in cross-model detection tasks when compared to existing state-of-the-art techniques. Additionally, the proposed method demonstrates resilience against various types of real-world image noise perturbations such as compression and blurring. Extensive ablation studies further corroborate the synergistic benefits of fusing multi-model forensic features with frequency-aware learning, underscoring the efficacy of our approach.

Subject: Graphics

Publish: 2025-04-02 03:57:12 UTC


#4 Real Time Animator: High-Quality Cartoon Style Transfer in 6 Animation Styles on Images and Videos [PDF] [Copy] [Kimi] [REL]

Authors: Liuxin Yang, Priyanka Ladha

This paper presents a comprehensive pipeline that integrates state-of-the-art techniques to achieve high-quality cartoon style transfer for educational images and videos. The proposed approach combines the Inversion-based Style Transfer (InST) framework for both image and video style stylization, the Pre-Trained Image Processing Transformer (IPT) for post-denoising, and the Domain-Calibrated Translation Network (DCT-Net) for more consistent video style transfer. By fine-tuning InST with specific cartoon styles, applying IPT for artifact reduction, and leveraging DCT-Net for temporal consistency, the pipeline generates visually appealing and educationally effective stylized content. Extensive experiments and evaluations using the scenery and monuments dataset demonstrate the superiority of the proposed approach in terms of style transfer accuracy, content preservation, and visual quality compared to the baseline method, AdaAttN. The CLIP similarity scores further validate the effectiveness of InST in capturing style attributes while maintaining semantic content. The proposed pipeline streamlines the creation of engaging educational content, empowering educators and content creators to produce visually captivating and informative materials efficiently.

Subject: Graphics

Publish: 2025-04-01 23:56:11 UTC


#5 ConfEviSurrogate: A Conformalized Evidential Surrogate Model for Uncertainty Quantification [PDF] [Copy] [Kimi] [REL]

Authors: Yuhan Duan, Xin Zhao, Neng Shi, Han-Wei Shen

Surrogate models, crucial for approximating complex simulation data across sciences, inherently carry uncertainties that range from simulation noise to model prediction errors. Without rigorous uncertainty quantification, predictions become unreliable and hence hinder analysis. While methods like Monte Carlo dropout and ensemble models exist, they are often costly, fail to isolate uncertainty types, and lack guaranteed coverage in prediction intervals. To address this, we introduce ConfEviSurrogate, a novel Conformalized Evidential Surrogate Model that can efficiently learn high-order evidential distributions, directly predict simulation outcomes, separate uncertainty sources, and provide prediction intervals. A conformal prediction-based calibration step further enhances interval reliability to ensure coverage and improve efficiency. Our ConfEviSurrogate demonstrates accurate predictions and robust uncertainty estimates in diverse simulations, including cosmology, ocean dynamics, and fluid dynamics.

Subjects: Machine Learning , Graphics , Machine Learning

Publish: 2025-04-03 15:44:14 UTC


#6 DualMS: Implicit Dual-Channel Minimal Surface Optimization for Heat Exchanger Design [PDF] [Copy] [Kimi] [REL]

Authors: Weizheng Zhang, Hao Pan, Lin Lu, Xiaowei Duan, Xin Yan, Ruonan Wang, Qiang Du

Heat exchangers are critical components in a wide range of engineering applications, from energy systems to chemical processing, where efficient thermal management is essential. The design objectives for heat exchangers include maximizing the heat exchange rate while minimizing the pressure drop, requiring both a large interface area and a smooth internal structure. State-of-the-art designs, such as triply periodic minimal surfaces (TPMS), have proven effective in optimizing heat exchange efficiency. However, TPMS designs are constrained by predefined mathematical equations, limiting their adaptability to freeform boundary shapes. Additionally, TPMS structures do not inherently control flow directions, which can lead to flow stagnation and undesirable pressure drops. This paper presents DualMS, a novel computational framework for optimizing dual-channel minimal surfaces specifically for heat exchanger designs in freeform shapes. To the best of our knowledge, this is the first attempt to directly optimize minimal surfaces for two-fluid heat exchangers, rather than relying on TPMS. Our approach formulates the heat exchange maximization problem as a constrained connected maximum cut problem on a graph, with flow constraints guiding the optimization process. To address undesirable pressure drops, we model the minimal surface as a classification boundary separating the two fluids, incorporating an additional regularization term for area minimization. We employ a neural network that maps spatial points to binary flow types, enabling it to classify flow skeletons and automatically determine the surface boundary. DualMS demonstrates greater flexibility in surface topology compared to TPMS and achieves superior thermal performance, with lower pressure drops while maintaining a similar heat exchange rate under the same material cost.

Subjects: Optimization and Control , Graphics , Machine Learning

Publish: 2025-03-02 15:24:38 UTC