---
title: High-Fidelity and Real-Time Novel View Synthesis for Dynamic Scenes
url: https://www.ml-quant.com/papers/arxiv/2310.08585/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2310.08585
source_url: https://arxiv.org/pdf/2310.08585.pdf
featured: 2023-10-16
citations: 76
topic: Other
---


# High-Fidelity and Real-Time Novel View Synthesis for Dynamic Scenes

Real-Time View Synthesis: Im4D is a hybrid scene representation that combines grid-based geometry and multi-view image-based appearance for dynamic view synthesis from multi-view videos, providing high-quality rendering and efficient training.

- Source: https://arxiv.org/pdf/2310.08585.pdf
- Identifier: arXiv:2310.08585
- Released: 2023-10-12
- First featured: Quant Letter No. 21 (2023-10-16): https://www.ml-quant.com/issues/2023-10-16/
- Citations (Semantic Scholar): 76
- Published in: SIGGRAPH Asia 2023 Conference Papers
- Topic: Other

## Related

- [Depth Anything V2](https://www.ml-quant.com/papers/arxiv/2406.09414/): Depth Anything V2 is a new model for monocular depth estimation, using synthetic and large-scale pseudo-labeled real images for faster, more accurate results and setting a new evaluation benchmark.
- [MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark](https://www.ml-quant.com/papers/arxiv/2406.01574/): MMLU-Pro, an improved dataset, expands the Massive Multitask Language Understanding benchmark by adding tougher questions and more choices, serving as a better benchmark to monitor progress in the field.
- [Qwen2.5-Coder Technical Report](https://www.ml-quant.com/papers/arxiv/2409.12186/): The report unveils the Qwen2.5-Coder series, an improvement from its predecessor, showcasing remarkable code generation abilities and achieving top-tier performance in various code-related tasks.
- [Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian Splatting](https://www.ml-quant.com/papers/arxiv/2310.10642/): The 4DGS model is introduced, capable of reconstructing dynamic 3D scenes from 2D images and generating diverse views over time, providing real-time rendering efficiency.
- [Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives](https://www.ml-quant.com/papers/arxiv/2311.18259/): Understanding Human Activity: The paper presents Ego-Exo4D, a large-scale video dataset and benchmark challenge featuring human activities from various perspectives, aimed at improving first-person video understanding.
- [FateZero: Fusing Attentions for Zero-shot Text-based Video Editing](https://www.ml-quant.com/papers/arxiv/2303.09535/): Text-based Video Editing: The article introduces FateZero, a new method for editing real-world videos using text, which outperforms previous models in maintaining video structure, motion, and frame consistency.
