본 연구에서는 데이터 품질 향상을 목표로, 혼합형 데이터에 대한 균형잡힌 증강을 위한 BAMT-GAN(Balanced Augmentation for Mixed Tabular GAN)이라는 새로운 데이터 증강 기법을 제안한다. BAMT-GAN은 생성 ...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=A110230833
2023
Korean
데이터 증강 ; 테이블 데이터 ; 학습데이터 ; 생성 모델 ; 딥러닝 ; Data Augmentation ; Tabular Data ; Training Data ; Generative Model ; GAN ; Deep Learning
004
학술저널
1-22(22쪽)
0
상세조회0
다운로드본 연구에서는 데이터 품질 향상을 목표로, 혼합형 데이터에 대한 균형잡힌 증강을 위한 BAMT-GAN(Balanced Augmentation for Mixed Tabular GAN)이라는 새로운 데이터 증강 기법을 제안한다. BAMT-GAN은 생성 ...
본 연구에서는 데이터 품질 향상을 목표로, 혼합형 데이터에 대한 균형잡힌 증강을 위한 BAMT-GAN(Balanced Augmentation for Mixed Tabular GAN)이라는 새로운 데이터 증강 기법을 제안한다. BAMT-GAN은 생성 모델, 클러스터링, 오버샘플링 기법을 결합하여 데이터 불균형 문제를 해결하며, 이를 통해 데이터의 다양성을 증가시키고 결과적으로 생성된 예측 모델의 정확도를 향상시킨다. 이 기법은 유사한 특성을 가진 데이터를 클러스터링하고, CNN 기반 GAN 모델을 이용해 새로운 데이터를 생성하는 과정을 포함한다. 다양한 분류 알고리즘을 통합하여 정확한 예측값을 도출하고, 층화 추출법을 이용해 균형 잡힌 데이터셋을 최종적으로 구성한다. 제안 기법은 의료, 마케팅, 금융, 환경 도메인에서 생성된 다양한 데이터셋을 활용하여 실험을 수행하였으며, 이를 통해 BAMT-GAN이 기존 데이터 증강 기법에 비해 데이터 불균형을 효과적으로 해결하고, 예측모델의 성능을 상당히 향상시킴을 확인하였다. 즉 BAMT-GAN을 활용한 데이터 증강 작업은 고신뢰도 예측모델을 구축하는데 필요한 고품질 학습데이터를 생성할 수 있게 한다.
다국어 초록 (Multilingual Abstract)
In this study, we present BAMT-GAN (Balanced Augmentation for Mixed Tabular GAN), a novel data augmentation technique aimed at enhancing the quality of datasets and addressing the issues of data imbalance in mixed tabular data. BAMT-GAN effectively in...
In this study, we present BAMT-GAN (Balanced Augmentation for Mixed Tabular GAN), a novel data augmentation technique aimed at enhancing the quality of datasets and addressing the issues of data imbalance in mixed tabular data. BAMT-GAN effectively integrates generative models, clustering, and oversampling techniques to alleviate data imbalance issues, thus increasing data diversity and enhancing the accuracy of the resulting predictive models. This technique involves clustering data with similar characteristics and generating new data using a CNN-based GAN model. It amalgamates various classification algorithms to derive accurate predictions and employs stratified sampling to ultimately construct a balanced dataset. We conducted experiments utilizing diverse datasets generated from the domains of healthcare, marketing, finance, and environmental science. Through these experiments, we have ascertained that BAMT-GAN significantly mitigates data imbalance and notably enhances the performance of predictive models compared to traditional data augmentation techniques. This suggests that the quality of training data essential for constructing reliable predictive models can be considerably improved through data augmentation operations utilizing BAMT-GAN.
목차 (Table of Contents)
프롬프트 엔지니어링과 BERTopic을 활용한 고령자 교통사고 인식 분석 및 개선 방향 제시
지식관리시스템 활용 지원이 기술수용 및 정보제공 행동에 미치는 영향: 상호 피드백과 협력 지향성의 역할
디지털 전환 사회의 신고령층 디지털 리터러시 측정 도구 개발
반도체 팹의 OHT 네트워크에서 딥러닝 기반의 섹션 별 단기간 정체수준 예측 모델