통합 검색 | Korea Science

Convolutional GRU and Attention based Fall Detection Integrating with Human Body Keypoints and DensePose

Yi Zheng;Cunyi Liao;Ruifeng Xiao;Qiang He
- KSII Transactions on Internet and Information Systems (TIIS)
- /
- 제18권9호
- /
- pp.2782-2804
- /
- 2024
The integration of artificial intelligence technology with medicine has rapidly evolved, with increasing demands for quality of life. However, falls remain a significant risk leading to severe injuries and fatalities, especially among the elderly. Therefore, the development and application of computer vision-based fall detection technologies have become increasingly important. In this paper, firstly, the keypoint detection algorithm ViTPose++ is used to obtain the coordinates of human body keypoints from the camera images. Human skeletal feature maps are generated from this keypoint coordinate information. Meanwhile, human dense feature maps are produced based on the DensePose algorithm. Then, these two types of feature maps are confused as dual-channel inputs for the model. The convolutional gated recurrent unit is introduced to extract the frame-to-frame relevance in the process of falling. To further integrate features across three dimensions (spatio-temporal-channel), a dual-channel fall detection algorithm based on video streams is proposed by combining the Convolutional Block Attention Module (CBAM) with the ConvGRU. Finally, experiments on the public UR Fall Detection Dataset demonstrate that the improved ConvGRU-CBAM achieves an F1 score of 92.86% and an AUC of 95.34%.
https://doi.org/10.3837/tiis.2024.09.016 인용 PDF HTML

Enhancing 3D Excavator Pose Estimation through Realism-Centric Image Synthetization and Labeling Technique

Tianyu Liang;Hongyang Zhao;Seyedeh Fatemeh Saffari;Daeho Kim
- 국제학술발표논문집
- /
- The 10th International Conference on Construction Engineering and Project Management
- /
- pp.1065-1072
- /
- 2024
Previous approaches to 3D excavator pose estimation via synthetic data training utilized a single virtual excavator model, low polygon objects, relatively poor textures, and few background objects, which led to reduced accuracy when the resulting models were tested on differing excavator types and more complex backgrounds. To address these limitations, the authors present a realism-centric synthetization and labeling approach that synthesizes results with improved image quality, more detailed excavator models, additional excavator types, and complex background conditions. Additionally, the data generated includes dense pose labels and depth maps for the excavator models. Utilizing the realism-centric generation method, the authors achieved significantly greater image detail, excavator variety, and background complexity for potentially improved labeling accuracy. The dense pose labels, featuring fifty points instead of the conventional four to six, could allow inferences to be made from unclear excavator pose estimates. The synthesized depth maps could be utilized in a variety of DNN applications, including multi-modal data integration and object detection. Our next step involves training and testing DNN models that would quantify the degree of accuracy enhancement achieved by increased image quality, excavator diversity, and background complexity, helping lay the groundwork for broader application of synthetic models in construction robotics and automated project management.
https://doi.org/10.6106/ICCEPM.2024.1065 인용 PDF

자율 수중 로봇을 위한 사실적인 실시간 고밀도 3차원 Mesh 지도 작성 (Photorealistic Real-Time Dense 3D Mesh Mapping for AUV)

이정우;조영근
- 로봇학회논문지
- /
- 제19권2호
- /
- pp.188-195
- /
- 2024
This paper proposes a photorealistic real-time dense 3D mapping system that utilizes a neural network-based image enhancement method and mesh-based map representation. Due to the characteristics of the underwater environment, where problems such as hazing and low contrast occur, it is hard to apply conventional simultaneous localization and mapping (SLAM) methods. At the same time, the behavior of Autonomous Underwater Vehicle (AUV) is computationally constrained. In this paper, we utilize a neural network-based image enhancement method to improve pose estimation and mapping quality and apply a sliding window-based mesh expansion method to enable lightweight, fast, and photorealistic mapping. To validate our results, we utilize real-world and indoor synthetic datasets. We performed qualitative validation with the real-world dataset and quantitative validation by modeling images from the indoor synthetic dataset as underwater scenes.
https://doi.org/10.7746/jkros.2024.19.2.188 인용 PDF

Onboard dynamic RGB-D simultaneous localization and mapping for mobile robot navigation

Canovas, Bruce;Negre, Amaury;Rombaut, Michele
- ETRI Journal
- /
- 제43권4호
- /
- pp.617-629
- /
- 2021
Although the actual visual simultaneous localization and mapping (SLAM) algorithms provide highly accurate tracking and mapping, most algorithms are too heavy to run live on embedded devices. In addition, the maps they produce are often unsuitable for path planning. To mitigate these issues, we propose a completely closed-loop online dense RGB-D SLAM algorithm targeting autonomous indoor mobile robot navigation tasks. The proposed algorithm runs live on an NVIDIA Jetson board embedded on a two-wheel differential-drive robot. It exhibits lightweight three-dimensional mapping, room-scale consistency, accurate pose tracking, and robustness to moving objects. Further, we introduce a navigation strategy based on the proposed algorithm. Experimental results demonstrate the robustness of the proposed SLAM algorithm, its computational efficiency, and its benefits for on-the-fly navigation while mapping.
https://doi.org/10.4218/etrij.2021-0061 인용 PDF KSCI

반자동적인 대응점 찾기를 이용한 3차원 얼굴 모델 생성 (Building a 3D Morphable Face Model using Finding Semi-automatic Dense Correspondence)

최인호;조선영;김대진
- 한국정보과학회논문지:컴퓨팅의 실제 및 레터
- /
- 제14권7호
- /
- pp.723-727
- /
- 2008
2D 기반의 얼굴 분석 및 처리 알고리즘은 포즈 및 조명에 강인하지 못한 문제점들이 존재한다. 이러한 이유로 과거 3D 기반의 얼굴 분석 및 처리 분야에 많은 연구를 진행하려 하였지만, 컴퓨팅 파워의 한계와 고속 스캐너의 부재 등으로 많은 연구가 진행되지 못하였다. 하지만 오늘날 하루가 다르게 빨라지고 있는 컴퓨터의 성능으로 인해 주춤했던 연구들이 다시 진행되고 있다. 이에 본 논문에서는 널리 알려진 선형 모델 기반의 3D morphable face model을 제작하고 성능을 높이는 방법에 대한 구현 및 dense correspondence 문제를 해결하기 위한 방법을 제안한다.
PDF KSCI

Keypoints-Based 2D Virtual Try-on Network System

Pham, Duy Lai;Ngyuen, Nhat Tan;Chung, Sun-Tae
- 한국멀티미디어학회논문지
- /
- 제23권2호
- /
- pp.186-203
- /
- 2020
Image-based Virtual Try-On Systems are among the most potential solution for virtual fitting which tries on a target clothes into a model person image and thus have attracted considerable research efforts. In many cases, current solutions for those fails in achieving naturally looking virtual fitted image where a target clothes is transferred into the body area of a model person of any shape and pose while keeping clothes context like texture, text, logo without distortion and artifacts. In this paper, we propose a new improved image-based virtual try-on network system based on keypoints, which we name as KP-VTON. The proposed KP-VTON first detects keypoints in the target clothes and reliably predicts keypoints in the clothes of a model person image by utilizing a dense human pose estimation. Then, through TPS transformation calculated by utilizing the keypoints as control points, the warped target clothes image, which is matched into the body area for wearing the target clothes, is obtained. Finally, a new try-on module adopting Attention U-Net is applied to handle more detailed synthesis of virtual fitted image. Extensive experiments on a well-known dataset show that the proposed KP-VTON performs better the state-of-the-art virtual try-on systems.
https://doi.org/10.9717/kmms.2020.23.2.186 인용 PDF KSCI HTML

UV-map 기반의 신경망 학습을 이용한 조립 설명서에서의 부품의 자세 추정 (UV Mapping Based Pose Estimation of Furniture Parts in Assembly Manuals)

강이삭;조남익
- 한국방송∙미디어공학회:학술대회논문집
- /
- 한국방송∙미디어공학회 2020년도 하계학술대회
- /
- pp.667-670
- /
- 2020
최근에는 증강현실, 로봇공학 등의 분야에서 객체의 위치 검출 이외에도, 객체의 자세에 대한 추정이 요구되고 있다. 객체의 자세 정보가 포함된 데이터셋은 위치 정보만 포함된 데이터셋에 비하여 상대적으로 매우 적기 때문에 인공 신경망 구조를 활용하기 어려운 측면이 있으나, 최근에 들어서는 기계학습 기반의 자세 추정 알고리즘들이 여럿 등장하고 있다. 본 논문에서는 이 가운데 Dense 6d Pose Object detector (DPOD) [11]의 구조를 기반으로 하여 가구의 조립 설명서에 그려진 가구 부품들의 자세를 추정하고자 한다. DPOD [11]는 입력으로 RGB 영상을 받으며, 해당 영상에서 자세를 추정하고자 하는 객체의 영역에 해당하는 픽셀들을 추정하고, 객체의 영역에 해당되는 각 픽셀에서 해당 객체의 3D 모델의 UV map 값을 추정한다. 이렇게 픽셀 개수만큼의 2D - 3D 대응이 생성된 이후에는, RANSAC과 PnP 알고리즘을 통해 RGB 영상에서의 객체와 객체의 3D 모델 간의 변환 관계 행렬이 구해지게 된다. 본 논문에서는 사전에 정해진 24개의 자세 후보들을 기반으로 가구 부품의 3D 모델을 2D에 투영한 RGB 영상들로 인공 신경망을 학습하였으며, 평가 시에는 실제 조립 설명서에서의 가구 부품의 자세를 추정하였다. 실험 결과 IKEA의 Stefan 의자 조립 설명서에 대하여 100%의 ADD score를 얻었으며, 추정 자세가 자세 후보군 중 정답 자세에 가장 근접한 경우를 정답으로 평가했을 때 100%의 정답률을 얻었다. 제안하는 신경망을 사용하였을 때, 가구 조립 설명서에서 가구 부품의 위치를 찾는 객체 검출기(object detection network)와, 각 개체의 종류를 구분하는 객체 리트리벌 네트워크(retrieval network)를 함께 사용하여 최종적으로 가구 부품의 자세를 추정할 수 있다.
PDF

Transfer Learning Models for Enhanced Prediction of Cracked Tires

Candra Zonyfar;Taek Lee;Jung-Been Lee;Jeong-Dong Kim
- Journal of Platform Technology
- /
- 제11권6호
- /
- pp.13-20
- /
- 2023
Regularly inspecting vehicle tires' condition is imperative for driving safety and comfort. Poorly maintained tires can pose fatal risks, leading to accidents. Unfortunately, manual tire visual inspections are often considered no less laborious than employing an automatic tire inspection system. Nevertheless, an automated tire inspection method can significantly enhance driver compliance and awareness, encouraging routine checks. Therefore, there is an urgency for automated tire inspection solutions. Here, we focus on developing a deep learning (DL) model to predict cracked tires. The main idea of this study is to demonstrate the comparative analysis of DenseNet121, VGG-19 and EfficientNet Convolution Neural Network-based (CNN) Transfer Learning (TL) and suggest which model is more recommended for cracked tire classification tasks. To measure the model's effectiveness, we experimented using a publicly accessible dataset of 1028 images categorized into two classes. Our experimental results obtain good performance in terms of accuracy, with 0.9515. This shows that the model is reliable even though it works on a dataset of tire images which are characterized by homogeneous color intensity.
PDF

춤추는 아바타: 당신도 싸이처럼 춤을 출 수 있다. (Dancing Avatar: You can dance like PSY too)

구동준;주영돈;브이 반 만;이정우;안희준
- 한국방송∙미디어공학회:학술대회논문집
- /
- 한국방송∙미디어공학회 2021년도 추계학술대회
- /
- pp.256-259
- /
- 2021
본 논문에서는 사람을 키넥트로 촬영하여 3 차원 아바타로 복원하여 연예인처럼 춤을 추게 하는 기술을 설계 구현하였다. 기존의 순수 딥러닝 기반 방식과 달리 본 기술은 3 차원 인체 모델을 사용하여 안정적이고 자유로운 결과를 얻을 수 있다. 우선 인체 모델의 기하학적 정보는 3 차원 조인트를 사용하여 추정하고 DensePose를 통하여 정교한 텍스쳐를 복원한다. 여기에 3 차원 포인트-클라우드와 ICP 매칭 기법을 사용하여 의상 모델 정보를 복원한다. 이렇게 확보한 신체 모델과 의상 모델을 사용한 아바타는 신체 모델의 rigged 특성을 그대로 유지함으로써 애니메이션에 적합하여 PSY 의 <강남스타일>과 같은 춤을 자연스럽게 표현하였다. 개선할 점으로 인체와 의류 부분의 좀 더 정확한 분할과 분할과정에서 발생할 수 있는 노이즈의 제거 등을 확인되었다.
PDF

3차원 손 특징을 이용한 손 동작 인식에 관한 연구 (A study on hand gesture recognition using 3D hand feature)

배철수
- 한국정보통신학회논문지
- /
- 제10권4호
- /
- pp.674-679
- /
- 2006
본 논문에서는 3차원 손 특징 데이터를 이용한 동작 인식 시스템을 제안하고자 한다. 제안된 시스템은 3차원 센서에 의해 조밀한 범위의 영상을 생성하여 손 동작에 대한 3차원 특징을 추출하여 손 동작을 분류한다. 또한 다양한 조명과 배경하에서의 손을 견실하게 분할하고 색상 정보와 상관이 없어 수화와 같은 복잡한 손 동작에 대해서도 견실한 인식능력을 나타낼 수가 있다. 제안된 방법의 전체적인 순서는 3차원 영상 획득, 팔 분할, 손과 팔목 분할, 손 자세 추정, 3차원 특징 추출, 그리고 동작 분류로 구성되어 있고, 수화 자세에 대한 인식 실험으로 제안된 시스템의 효율성을 입증하였다.
PDF KSCI

검색결과 16건 처리시간 0.025초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)