F114 - Informatique, Logiciel et Intelligence artificielle
Research institute :
R450 - Institut NUMEDIART pour les Technologies des Arts Numériques R300 - Institut de Recherche en Technologies de l'Information et Sciences de l'Informatique
Nguyen, D.T., Li, W., Ogunbona, P.O.: Human detection from images and videos: a survey. Pattern Recogn. 51, 148–175 (2016)
Zhang, S., Benenson, R., Omran, M., Hosang, J., Schiele, B.: How far are we from solving pedestrian detection? In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 1259–1267 (2016)
Leal-Taixé, L., Milan, A., Reid, I., Roth, S., Schindler, K.: Motchallenge 2015: towards a benchmark for multi-target tracking. arXiv preprint arXiv:1504.01942 (2015)
Newell, A., Yang, K., Deng, J.: Stacked hourglass networks for human pose estimation. In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, 11–14 October 2016, Proceedings, Part VIII 14, pp. 483–499. Springer (2016)
Xiao, B., Wu, H., Wei, Y.: Simple baselines for human pose estimation and tracking. In: Proceedings of the European Conference on Computer Vision (ECCV), pp. 466–481 (2018)
Pandya, S., et al.: Federated learning for smart cities: a comprehensive survey. Sustain. Energy Technol. Assessm. 55, 102987 (2023)
Gunasekaran, K.P., Jaiman, N.: Now you see me: robust approach to partial occlusions. arXiv preprint arXiv:2304.11779 (2023)
Wen, L., et al.: UA-DETRAC: a new benchmark and protocol for multi-object detection and tracking. Comput. Vis. Image Underst. 193, 102907 (2020)
Sharifani, K., Amini, M.: Machine learning and deep learning: a review of methods and applications. World Inf. Technol. Eng. J. 10(07), 3897–3904 (2023)
Ye, H., Zhao, J., Pan, Y., Cherr, W., He, L., Zhang, H.: Robot person following under partial occlusion. In: 2023 IEEE International Conference on Robotics and Automation (ICRA), pp. 7591–7597. IEEE (2023)
Ouardirhi, Z., Amel, O., Zbakh, M., Mahmoudi, S.A.: FuDensityNet: fusion-based density-enhanced network for occlusion handling. In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pp. 3008–3017 (2023). https://doi.org/10.5220/0012425400003660, https://www.insticc.org/node/TechnicalProgram/visigrapp/2024/presentationDetails124254
Steyaert, S., et al.: Multimodal data fusion for cancer biomarker discovery with deep learning. Nat. Mach. Intell. 5(4), 351–362 (2023)
Pérez-Hernández, F., Tabik, S., Lamas, A., Olmos, R., Fujita, H., Herrera, F.: Object detection binary classifiers methodology based on deep learning to identify small objects handled similarly: Application in video surveillance. Knowl.-Based Syst. 194, 105590 (2020)
Chen, X., Ma, H., Wan, J., Li, B., Xia, T.: Multi-view 3d object detection network for autonomous driving. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 1907–1915 (2017)
Ouardirhi, Z., Mahmoudi, S.A., Zbakh, M.: Enhancing object detection in smart video surveillance: a survey of occlusion-handling approaches. Electronics 13(3), 541 (2024)
Kortylewski, A., He, J., Liu, Q., Yuille, A.L.: Compositional convolutional neural networks: a deep architecture with innate robustness to partial occlusion. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 8940–8949 (2020)
Bagautdinov, T., Fleuret, F., Fua, P.: Probability occupancy maps for occluded depth images. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 2829–2837 (2015)
Sun, Y., Kortylewski, A., Yuille, A.: Amodal segmentation through out-of-task and out-of-distribution generalization with a bayesian model. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1215–1224 (2022)
Ali, W., Abdelkarim, S., Zidan, M., Zahran, M., El Sallab, A.: Yolo3d: end-to-end real-time 3d oriented object bounding box detection from lidar point cloud. In: Proceedings of the European Conference on Computer Vision (ECCV) Workshops (2018)
Yang, C., Ablavsky, V., Wang, K., Feng, Q., Betke, M.: Learning to separate: detecting heavily-occluded objects in urban scenes. In: European Conference on Computer Vision, pp. 530–546. Springer (2020)
Takahashi, M., Ji, Y., Umeda, K., Moro, A.: Expandable YOLO: 3d object detection from RGB-D images. In: 2020 21st International Conference on Research and Education in Mechatronics (REM), pp. 1–5. IEEE (2020)
Jenkins, P., et al.: CountNet3D: a 3D computer vision approach to infer counts of occluded objects. In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pp. 3008–3017 (2023)
Reynolds, J., Nagesh, C.K., Gurari, D.: Salient object detection for images taken by people with vision impairments. arXiv preprint arXiv:2301.05323 (2023)
Mirbod, O., Choi, D., Heinemann, P.H., Marini, R.P., He, L.: On-tree apple fruit size estimation using stereo vision with deep learning-based occlusion handling. Biosys. Eng. 226, 27–42 (2023)
Ouyang, W., et al.: Deepid-net: deformable deep convolutional neural networks for object detection. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 2403–2412 (2015)
Geiger, A., Lenz, P., Urtasun, R.: Are we ready for autonomous driving? the kitti vision benchmark suite. In: 2012 IEEE Conference on Computer Vision and Pattern Recognition, pp. 3354–3361. IEEE (2012)
Sharma, P., Gupta, S., Vyas, S., Shabaz, M.: Retracted: object detection and recognition using deep learning-based techniques. IET Commun. 17(13), 1589–1599 (2023)
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 770–778 (2016)
Sandler, M., Howard, A., Zhu, M., Zhmoginov, A., Chen, L.C.: Mobilenetv2: inverted residuals and linear bottlenecks. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 4510–4520 (2018)
Simonyan, K., Zisserman, A.: Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 (2014)
Lin, T.Y., Goyal, P., Girshick, R., He, K., Dollár, P.: Focal loss for dense object detection. In: Proceedings of the IEEE International Conference on Computer Vision, pp. 2980–2988 (2017)
Liu, W., et al.: SSD: single shot multibox detector. In: European Conference on Computer Vision, pp. 21–37. Springer (2016)
Sozzi, M., Cantalamessa, S., Cogato, A., Kayad, A., Marinello, F.: Automatic bunch detection in white grape varieties using YOLOv3, YOLOv4, and YOLOv5 deep learning algorithms. Agronomy 12(2), 319 (2022)
Li, C., et al.: YOLOv6: a single-stage object detection framework for industrial applications. arXiv preprint arXiv:2209.02976 (2022)
Wang, C.Y., Bochkovskiy, A., Liao, H.Y.M.: YOLOv7: trainable bag-of-freebies sets new state-of-the-art for real-time object detectors. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 7464–7475 (2023)
Huang, Z., Li, L., Krizek, G.C., Sun, L.: Research on traffic sign detection based on improved YOLOv8. J. Comput. Commun. 11(7), 226–232 (2023)
Kortylewski, A., Liu, Q., Wang, H., Zhang, Z., Yuille, A.: Combining compositional models and deep networks for robust object classification under occlusion. In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pp. 1333–1341 (2020)