Empirical study of feature extraction approaches for image captioning in Vietnamese
Tác giả: Khang Nguyen
Số trang:
P. 327-346
Tên tạp chí:
Tin học & Điều khiển học
Số phát hành:
V.38-N.4
Kiểu tài liệu:
Tạp chí trong nước
Nơi lưu trữ:
03 Quang Trung
Mã phân loại:
005
Ngôn ngữ:
Tiếng Anh
Từ khóa:
Grid features, region features, image captioning, Viecap4h, uit-viic, faster R-CNN, cascade R-CNN, grid R-CNN, Vinvl
Chủ đề:
Computer science
Tóm tắt:
This study focus on the image captioning problem in Vietnamese. Indetail, an empirical study of grid-based and region-based feature extraction approaches using currentstate-of-the-art object detection methods is investigated to explore the suitable way to represent theimages in the model space. Each feature type represents images, and the image captioning task istrained using the Transformer-based model. The effectiveness of different feature types is exploredon two Vietnamese datasets: UIT-ViIC and VieCap4H, the two standard benchmark datasets. Theexperimental results show crucial insight into the feature extraction task for image captioning inVietnamese.
Tạp chí liên quan
- Kết quả điều trị nhắm trúng đích ở bệnh nhân ung thư phổi không tế bào nhỏ có đột biến phức hợp trên gen EGFR tại Bệnh viện Phổi Trung ương
- Kết quả bước đầu của phương pháp cắt tách dưới niêm mạc điều trị tổn thương ung thư sớm và tiền ung thư đường tiêu hóa tại Trung tâm Nội soi – Bệnh viện Đại học Y Hà Nội
- Một số yếu tố liên quan đến thời gian nằm viện của người bệnh rối loạn lo âu lan toả điều trị nội trú tại viện sức khoẻ tâm thần
- Đánh giá kết quả kết hợp xương bằng nẹp vít khoá điều trị gãy kín thân xương cánh tay tại Bệnh viện Đa khoa tỉnh Thanh Hóa
- Giá trị của Interleukin-6 huyết thanh trong dự báo biến cố nội viện ở bệnh nhân suy tim mất bù cấp





