GitHub - mhseo10/2022-SW-AI-Contest

customocr

경로 설정 방법: `customocr`의 하위 경로에 data 파일 위치

. customocr
├── font
├── hallymocr
├── ocr.ipynb
├── ...
├── open ==> dacon 파일
├── word_h ==> 구글 드라이브 파일
├── word_v
├── ko.txt
└── train.txt

valid_data 조절 방법(학습 시간 단축)

OCR 모델 생성 전 validation data 생성 코드 추가('train/' 으로 가정)

import pandas as pd

###### train.csv를 이용해 개수 조절
train_csv_path = 'open/train.csv'
train_csv = pd.read_csv(train_csv_path)
train_csv[:10].to_csv('open/train.txt', sep='\t', header=False, index=False)


###### lmdb 파일 생성, 운영체제에 맞게 주석 해제
# # if window
# !python ./hallymocr/create_lmdb_dataset.py --inputPath ./open/ --gtFile ./open/train.txt --outputPath ./result/train --file_size <전체 데이터 크기(GB)>
# # if linux
# !python3 ./hallymocr/create_lmdb_dataset.py --inputPath ./open/ --gtFile ./open/train.txt --outputPath ./result/train --file_size <전체 데이터 크기(GB)>

opt['valid_data'] 수정

opt = {
    'exp_name': None,
    'train_data': './result/',
    'valid_data': './result/train',
    ...
}

Name		Name	Last commit message	Last commit date
Latest commit History 47 Commits
aihub		aihub
hallymocr		hallymocr
.gitignore		.gitignore
README.md		README.md
change_label.ipynb		change_label.ipynb
create imdb txtfile.ipynb		create imdb txtfile.ipynb
ocr.ipynb		ocr.ipynb
requirements.yaml		requirements.yaml

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

customocr

경로 설정 방법: `customocr`의 하위 경로에 data 파일 위치

valid_data 조절 방법(학습 시간 단축)

- json_file 컨버터

About

Releases

Packages

Contributors 4

Languages

mhseo10/2022-SW-AI-Contest

Folders and files

Latest commit

History

Repository files navigation

customocr

경로 설정 방법: customocr의 하위 경로에 data 파일 위치

valid_data 조절 방법(학습 시간 단축)

- json_file 컨버터

About

Resources

Stars

Watchers

Forks

Releases

Packages 0

Contributors 4

Languages

경로 설정 방법: `customocr`의 하위 경로에 data 파일 위치

Packages