sphong

Assistant professor Department of Molecular Biology Jeonbuk National University

Seungpyo Hong

SNUH data curation

요약 임상 데이터의 체계적 수집 및 관리를 위해 데이터를 재구성하였습니다. 1개의 표에 모든 정보를 기록하고 있던 기존 데이터 수집 체계를 5개의 excel 파일, 18개의 table로 분리하였고, 기존 자료 (2024년 6월 28일 기준) 내용을 입력하였습니다. 비슷한 의미를 지닌 단위를 데이터를 분류하여, 각각의 데이터를 쉽게 관리하는 것이 목적입니다. 데이터 관리 측면에서는 출력 (fetch), 입력 (insert), 수정 (update), […]

SNUH data curation Read More »

Understanding the Concept of Statistics Through Examples

Library loading and functions RNA expression and protein expression Measure of the concentration of the protein X Interepretation A Statistic: A Single Number that Represents the Data Simple Evaluation Random Incidences 0.5148158483005283-0.68221888442359540.43813562081717283-0.24953267639975260.10300182994047037 Statistical Inference P-value Statistical inference

Understanding the Concept of Statistics Through Examples Read More »

Mid-term project – Deep Learning and Drug Discovery

1. 소개 본 과제는 신약 개발 과정 중 hit discovery 과정에 대해 복습하고, 실제 연구에서 데이터와 deep learning 기반 방법이 어떻게 활용될 수 있는지를 이해하는 것을 목표로 합니다. 특히, 데이터를 수집하고 분석하며, computational 및 AI 기반 방법을 이용해 문제를 해결하는 과정을 직접 경험함으로써 데이터 기반 연구(data-driven research)의 전체 흐름을 이해하는 데 목적이 있습니다. 학생들은 과제

Mid-term project – Deep Learning and Drug Discovery Read More »

Pattern – Bioinformatics

Introduction Pattern은 반복되는 무늬나 서열을 의미한다. DNA 서열과 단백질 서열에는 다양한 종류의 pattern이 존재한다. 생체에서 이러한 패턴이 나타나는 이유는 기능과 밀접히 연관되어 있다. 동일한 transcription factor (TF) 에 의해 발현이 조절되는 유전자의 promoter에는 해당 TF가 인지하고 결합할 수 있는 특정 DNA 서열이 있다. 유사한 기능을 하는 단백질의 경우 해당 기능 수행을 위한 부위가 잘 보존되어

Pattern – Bioinformatics Read More »

Activity for Phylogenetic Tree

UPGMA는 종이나 유전자 서열 간의 관계를 계통수로 표현하는 방법이다. 이 활동에서는 해당 알고리즘을 Python 코드로 구현하는 방법을 이해해본다. 또한, BioPython 라이브러리의 UPGMA 및 Neighbor-Joining(NJ) 알고리즘을 이용하여 계통수를 생성하는 과정을 실습한다. 1. UPGMA Unweighted Pair Group Method with Arithmetic Means (UPGMA)는 서열 간 distance를 이용해 phylogenetic tree를 구성하는 가장 기본적인 방법이다. 본 실습에서는 해당 알고리즘을 어떻게

Activity for Phylogenetic Tree Read More »

Scroll to Top