GENERATIV SUNIY INTELLEKT YORDAMIDA 2D GRAFIK TASVIRLARDAN 3D OBYEKTLARNI HOSIL QILISHNING USUL VA ALGORITMLARI
DOI:
https://doi.org/10.65164/s64ebf91Kalit so‘zlar:
2020 yilda Mildenhall va boshqalar tomonidan taklif etilgan Neural Radiance Field (NeRF) arxitekturasi 3D qayta tiklash sohasida inqilob yasadiAbstrak
Soʻnggi yillarda sunʼy intellekt va chuqur oʻrganish (deep learning) texnologiyalarining rivojlanishi natijasida 2D tasvirlardan 3D obʻektlarni qayta tiklash masalasi kompyuter koʻrishi va kompyuter grafikasidagi eng dolzarb yoʻnalishlardan biriga aylandi. Anʻanaviy fotogrammetrik usullar yuzlab rasmlarni talab qilsa, generativ modellar bitta yoki bir nechta 2D tasvirdan toʻliq uch oʻchamli shaklni hosil qila oladi [1, 2].
Havolalar
1. Artikova M.A., Bahromov A.A., Jo‘raboyev. F.A., (2024). 3D texnologiyalarni talimdagi
o‘rni. Yevroosiyo Akademik Tadqiqotlar Jurnalining Universal Indeks Kutubxonasi, 4(12
Special Issue), 727–729. https://lib.uniconflix.com/index.php/uilejar/article/view/44219
2. Bahromov, A., Jo‘raboyev, . F. (2024). Virtual reallikni sohalarda qo‘llanilishi.
Yevroosiyo Akademik Tadqiqotlar Jurnalining Universal Indeks Kutubxonasi, 4(12
Special Issue), 730–734. https://lib.uniconflix.com/index.php/uilejar/article/view/44220
3. Mildenhall B., Srinivasan P.P., Tancik M., Barron J.T., Ramamoorthi R., Ng R. (2020).
NeRF: Representing scenes as neural radiance fields for view synthesis. In ECCV 2020
(pp. 405–421). Springer. https://doi.org/10.1007/978-3-030-58452-8_24
4. Goodfellow I., Pouget-Abadie J., Mirza M., Xu B., Warde-Farley D., Ozair S., Courville
A., Bengio Y. (2014). Generative adversarial nets. Advances in Neural Information
Processing Systems, 27, 2672–2680.15
5. Rombach R., Blattmann A., Lorenz D., Esser P., Ommer B. (2022). High-resolution image
synthesis with latent diffusion models. In CVPR 2022 (pp. 10684–10695).
https://doi.org/10.1109/CVPR52688.2022.01042
6. Poole B., Jain A., Barron J.T., Mildenhall B. (2023). DreamFusion: Text-to-3D using 2D
diffusion. In ICLR 2023. https://arxiv.org/abs/2209.14988
7. Grand View Research. (2024). 3D Content Creation Market Size, Share & Trends Analysis
Report. San Francisco: GVR.
8. Hartley R., Zisserman A. (2004). Multiple View Geometry in Computer Vision (2nd ed.).
Cambridge University Press. https://doi.org/10.1017/CBO9780511811685
9. Pollefeys M., Nistér D., Frahm J.M., Akbarzadeh A., Mordohai P., Clipp B., ... & Lhuillier
M. (2008). Detailed real-time urban 3D reconstruction from video. IJCV, 78(2), 143–167.
https://doi.org/10.1007/s11263-007-0086-4
10. Müller T., Evans A., Schied C., Keller A. (2022). Instant neural graphics primitives with a
multiresolution hash encoding. ACM TOG, 41(4), 1–15.
https://doi.org/10.1145/3528223.3530127
11. Chan E.R., Lin C.Z., Chan M.A., Nagano K., Pan B., De Mello S., ... & Wetzstein G.
(2022). Efficient geometry-aware 3D generative adversarial networks. In CVPR 2022 (pp.
16123–16133). https://doi.org/10.1109/CVPR52688.2022.01565
12. Schwarz K., Liao Y., Niemeyer M., Geiger A. (2020). GRAF: Generative radiance fields
for 3D-aware image synthesis. NeurIPS 2020, 33, 20154–20166.
13. Niemeyer M., Geiger A. (2021). GIRAFFE: Representing scenes as compositional
generative neural feature fields. In CVPR 2021 (pp. 11453–11464).
https://doi.org/10.1109/CVPR46437.2021.01129
14. Ho J., Jain A., & Abbeel P. (2020). Denoising diffusion probabilistic models. NeurIPS
2020, 33, 6840–6851.
15. Liu R., Wu R., Van Hoorick B., Tokmakov P., Zakharov S., & Vondrick C. (2023). Zero-
1-to-3: Zero-shot one image to 3D object. In ICCV 2023 (pp. 9298–9309).
https://doi.org/10.1109/ICCV51070.2023.00853
16. Hong Y., Zhang K., Gu J., Bi S., Zhou Y., Liu D., ... & Xu Z. (2024). LRM: Large
reconstruction model for single image to 3D. In ICLR 2024.
https://arxiv.org/abs/2311.04400
17. Shi R., Chen H., Zhang Z., Liu M., Xu C., Wei X., ... & Zhang G. (2023). Zero123++: A
single image to consistent multi-view diffusion base model. arxiv:2310.15110.
https://arxiv.org/abs/2310.15110
18. Ranftl R., Bochkovskiy A., Koltun V. (2021). Vision transformers for dense prediction. In
ICCV 2021 (pp. 12179–12188). https://doi.org/10.1109/ICCV48922.2021.01196
19. Ranftl R., Lasinger K., Hafner D., Schindler K., Koltun V. (2020). Towards robust
monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer. IEEE
TPAMI, 44(3), 1623–1637. https://doi.org/10.1109/TPAMI.2020.3019967
20. Long X., Lin C., Wang P., Komura T., Wang W. (2022). SparseNeuS: Fast generalizable
neural surface reconstruction from sparse views. In ECCV 2022 (pp. 710–726). Springer.
https://doi.org/10.1007/978-3-031-20086-1_41
21. Bahromov A., Jo‘raboyev F. (2024). Virtual reallik texnologiyalarining matematik
modellarini tahlil qilish. Yevroosiyo Akademik Tadqiqotlar Jurnalining Universal Indeks
Kutubxonasi, 4 (12 Special Issue), 811–813.16
https://lib.uniconflix.com/index.php/uilejar/article/view/44243
22. Bahromov A.A., Ibodullayev S.N. (2022) Virtual reallik geometriyasi. "Экономика и
социум" №12(103)-2. 15-21.
https://cyberleninka.ru/article/n/virtual-reallik-geometriyasi-i/viewer
23. Yu A., Ye V., Tancik M., Kanazawa A. (2021). PixelNeRF: Neural radiance fields from
one or few images. In CVPR 2021 (pp. 4578–4587).
https://doi.org/10.1109/CVPR46437.2021.00455
24. Liu M., Xu C., Jin H., Chen L., Varma T.M., Xu Z., Su H. (2023). One-2-3-45: Any single
image to 3D mesh in 45 seconds without per-shape optimization. NeurIPS 2023, 36.
https://arxiv.org/abs/2306.16928
Fischer T., Gool L.V., Timofte R. (2024). Boosting novel view synthesis with geometric
priors. In WACV 2024. https://arxiv.org/abs/2310.09727
25. Verbin D., Hedman P., Mildenhall B., Zickler T., Barron J.T., Srinivasan P.P. (2022). RefNeRF: Structured view-dependent appearance for neural radiance fields. In CVPR 2022
(pp. 5491–5500). https://doi.org/10.1109/CVPR52688.2022.00541
26. Kerbl B., Kopanas G., Leimkühler T., Drettakis G. (2023). 3D Gaussian splatting for realtime radiance field rendering. ACM TOG, 42(4), 1–14. https://doi.org/10.1145/3592433
27. Chiu P., Zheng D., Hines J., Goldsmith D. (2023). Impact of 3D surgical planning on
operative outcomes in complex craniofacial procedures. Journal of Craniofacial Surgery,
34(2), 512–518. https://doi.org/10.1097/SCS.0000000000009201
28. Xu X., Tian Q., Lin J., Kumar S. (2023). Immersive product visualization: The effect of
3D models on e-commerce conversion rates. Journal of Retailing and Consumer Services,
72, 103274. https://doi.org/10.1016/j.jretconser.2023.103274
Weng J., Chen Y., Neven D., Yu X., Navarro-Alarcon D., Abbeel P. (2023). Neural grasp
distance fields for robot manipulation. In ICRA 2023 (pp. 1209–1215). IEEE.