GENERATIV SUNIY INTELLEKT YORDAMIDA 2D GRAFIK TASVIRLARDAN 3D OBYEKTLARNI HOSIL QILISHNING USUL VA ALGORITMLARI

Mualliflar

  • T.A.Xo‘jakulov Toshkent Amaliy fanlar universiteti, Kompyuter injiniringi kafedrasi, t.f.f.d (PhD), dotsent Muallif
  • A.A.Bahromov Toshkent Amaliy fanlar universiteti, Kompyuter injiniringi kafedrasi katta o‘qituvchisi Muallif
  • B.E.Boymurodov O‘zbekiston Jurnalistika va ommaviy kommunikatsiyalar universiteti, Mediadizayn kafedrasi katta o‘qituvchisi Muallif

DOI:

https://doi.org/10.65164/s64ebf91

Kalit so‘zlar:

2020 yilda Mildenhall va boshqalar tomonidan taklif etilgan Neural Radiance Field (NeRF) arxitekturasi 3D qayta tiklash sohasida inqilob yasadi

Abstrak

Soʻnggi yillarda sunʼy intellekt va chuqur oʻrganish (deep learning) texnologiyalarining rivojlanishi natijasida 2D tasvirlardan 3D obʻektlarni qayta tiklash masalasi kompyuter koʻrishi va kompyuter grafikasidagi eng dolzarb yoʻnalishlardan biriga aylandi. Anʻanaviy fotogrammetrik usullar yuzlab rasmlarni talab qilsa, generativ modellar bitta yoki bir nechta 2D tasvirdan toʻliq uch oʻchamli shaklni hosil qila oladi [1, 2]. 

Havolalar

1. Artikova M.A., Bahromov A.A., Jo‘raboyev. F.A., (2024). 3D texnologiyalarni talimdagi

o‘rni. Yevroosiyo Akademik Tadqiqotlar Jurnalining Universal Indeks Kutubxonasi, 4(12

Special Issue), 727–729. https://lib.uniconflix.com/index.php/uilejar/article/view/44219

2. Bahromov, A., Jo‘raboyev, . F. (2024). Virtual reallikni sohalarda qo‘llanilishi.

Yevroosiyo Akademik Tadqiqotlar Jurnalining Universal Indeks Kutubxonasi, 4(12

Special Issue), 730–734. https://lib.uniconflix.com/index.php/uilejar/article/view/44220

3. Mildenhall B., Srinivasan P.P., Tancik M., Barron J.T., Ramamoorthi R., Ng R. (2020).

NeRF: Representing scenes as neural radiance fields for view synthesis. In ECCV 2020

(pp. 405–421). Springer. https://doi.org/10.1007/978-3-030-58452-8_24

4. Goodfellow I., Pouget-Abadie J., Mirza M., Xu B., Warde-Farley D., Ozair S., Courville

A., Bengio Y. (2014). Generative adversarial nets. Advances in Neural Information

Processing Systems, 27, 2672–2680.15

5. Rombach R., Blattmann A., Lorenz D., Esser P., Ommer B. (2022). High-resolution image

synthesis with latent diffusion models. In CVPR 2022 (pp. 10684–10695).

https://doi.org/10.1109/CVPR52688.2022.01042

6. Poole B., Jain A., Barron J.T., Mildenhall B. (2023). DreamFusion: Text-to-3D using 2D

diffusion. In ICLR 2023. https://arxiv.org/abs/2209.14988

7. Grand View Research. (2024). 3D Content Creation Market Size, Share & Trends Analysis

Report. San Francisco: GVR.

8. Hartley R., Zisserman A. (2004). Multiple View Geometry in Computer Vision (2nd ed.).

Cambridge University Press. https://doi.org/10.1017/CBO9780511811685

9. Pollefeys M., Nistér D., Frahm J.M., Akbarzadeh A., Mordohai P., Clipp B., ... & Lhuillier

M. (2008). Detailed real-time urban 3D reconstruction from video. IJCV, 78(2), 143–167.

https://doi.org/10.1007/s11263-007-0086-4

10. Müller T., Evans A., Schied C., Keller A. (2022). Instant neural graphics primitives with a

multiresolution hash encoding. ACM TOG, 41(4), 1–15.

https://doi.org/10.1145/3528223.3530127

11. Chan E.R., Lin C.Z., Chan M.A., Nagano K., Pan B., De Mello S., ... & Wetzstein G.

(2022). Efficient geometry-aware 3D generative adversarial networks. In CVPR 2022 (pp.

16123–16133). https://doi.org/10.1109/CVPR52688.2022.01565

12. Schwarz K., Liao Y., Niemeyer M., Geiger A. (2020). GRAF: Generative radiance fields

for 3D-aware image synthesis. NeurIPS 2020, 33, 20154–20166.

13. Niemeyer M., Geiger A. (2021). GIRAFFE: Representing scenes as compositional

generative neural feature fields. In CVPR 2021 (pp. 11453–11464).

https://doi.org/10.1109/CVPR46437.2021.01129

14. Ho J., Jain A., & Abbeel P. (2020). Denoising diffusion probabilistic models. NeurIPS

2020, 33, 6840–6851.

15. Liu R., Wu R., Van Hoorick B., Tokmakov P., Zakharov S., & Vondrick C. (2023). Zero-

1-to-3: Zero-shot one image to 3D object. In ICCV 2023 (pp. 9298–9309).

https://doi.org/10.1109/ICCV51070.2023.00853

16. Hong Y., Zhang K., Gu J., Bi S., Zhou Y., Liu D., ... & Xu Z. (2024). LRM: Large

reconstruction model for single image to 3D. In ICLR 2024.

https://arxiv.org/abs/2311.04400

17. Shi R., Chen H., Zhang Z., Liu M., Xu C., Wei X., ... & Zhang G. (2023). Zero123++: A

single image to consistent multi-view diffusion base model. arxiv:2310.15110.

https://arxiv.org/abs/2310.15110

18. Ranftl R., Bochkovskiy A., Koltun V. (2021). Vision transformers for dense prediction. In

ICCV 2021 (pp. 12179–12188). https://doi.org/10.1109/ICCV48922.2021.01196

19. Ranftl R., Lasinger K., Hafner D., Schindler K., Koltun V. (2020). Towards robust

monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer. IEEE

TPAMI, 44(3), 1623–1637. https://doi.org/10.1109/TPAMI.2020.3019967

20. Long X., Lin C., Wang P., Komura T., Wang W. (2022). SparseNeuS: Fast generalizable

neural surface reconstruction from sparse views. In ECCV 2022 (pp. 710–726). Springer.

https://doi.org/10.1007/978-3-031-20086-1_41

21. Bahromov A., Jo‘raboyev F. (2024). Virtual reallik texnologiyalarining matematik

modellarini tahlil qilish. Yevroosiyo Akademik Tadqiqotlar Jurnalining Universal Indeks

Kutubxonasi, 4 (12 Special Issue), 811–813.16

https://lib.uniconflix.com/index.php/uilejar/article/view/44243

22. Bahromov A.A., Ibodullayev S.N. (2022) Virtual reallik geometriyasi. "Экономика и

социум" №12(103)-2. 15-21.

https://cyberleninka.ru/article/n/virtual-reallik-geometriyasi-i/viewer

23. Yu A., Ye V., Tancik M., Kanazawa A. (2021). PixelNeRF: Neural radiance fields from

one or few images. In CVPR 2021 (pp. 4578–4587).

https://doi.org/10.1109/CVPR46437.2021.00455

24. Liu M., Xu C., Jin H., Chen L., Varma T.M., Xu Z., Su H. (2023). One-2-3-45: Any single

image to 3D mesh in 45 seconds without per-shape optimization. NeurIPS 2023, 36.

https://arxiv.org/abs/2306.16928

Fischer T., Gool L.V., Timofte R. (2024). Boosting novel view synthesis with geometric

priors. In WACV 2024. https://arxiv.org/abs/2310.09727

25. Verbin D., Hedman P., Mildenhall B., Zickler T., Barron J.T., Srinivasan P.P. (2022). RefNeRF: Structured view-dependent appearance for neural radiance fields. In CVPR 2022

(pp. 5491–5500). https://doi.org/10.1109/CVPR52688.2022.00541

26. Kerbl B., Kopanas G., Leimkühler T., Drettakis G. (2023). 3D Gaussian splatting for realtime radiance field rendering. ACM TOG, 42(4), 1–14. https://doi.org/10.1145/3592433

27. Chiu P., Zheng D., Hines J., Goldsmith D. (2023). Impact of 3D surgical planning on

operative outcomes in complex craniofacial procedures. Journal of Craniofacial Surgery,

34(2), 512–518. https://doi.org/10.1097/SCS.0000000000009201

28. Xu X., Tian Q., Lin J., Kumar S. (2023). Immersive product visualization: The effect of

3D models on e-commerce conversion rates. Journal of Retailing and Consumer Services,

72, 103274. https://doi.org/10.1016/j.jretconser.2023.103274

Weng J., Chen Y., Neven D., Yu X., Navarro-Alarcon D., Abbeel P. (2023). Neural grasp

distance fields for robot manipulation. In ICRA 2023 (pp. 1209–1215). IEEE.

https://doi.org/10.1109/ICRA48891.2023.10160540

Yuklab olishlar

Nashr qilingan

2026-05-15