1K1 Rice Genome Project – Philippines (Overview)
📊 1kRG Activity Plan and Deliverables - IRRI
Table of Contents
- 🎯 Objectives
- 📅 Year 1 Targets
🎯 OBJECTIVES
- 🔧 Implement SNP-Seek Database Components of PH Rice GDB Portal at UPLB (including server configuration)
- 🌾 Populate Databases with information relevant to the Philippine TRVs from the public domain
- 🧬 Implement Bioinformatic Tools that will allow use of the new reference genomes and the re-sequencing data from the Philippine TRVs
- 📚 Training on Genome Assembly and Annotation
- 🖥️ Training on Database Development and Curation
- 🔍 Training on the Use of the PH Rice GDB Database and other associated tools
📅 Year 1 Targets
- 💻 Software Development and Server Setup on Y1 Q1–Q3
- 🚀 Pilot SNP-Seek Component of PH Rice GDB alpha version released by Y1 Q4
- 📊 Collated and Curated Passport and existing agro-morpho data for TRVs by Y1 Q3–Q4
- 🎓 First Training conducted by Y1 Q3
- 📚 Continuous Training of project staff by the IRRI team from Y1 Q1–Q3
- 👩🎓 1st Intake of Students as a special project in UPLB by Q2
📅 Quarter 1
📋 Activities
- Software Development and Server Setup
📅 Sub-Activities
- Evaluation and gap analysis of current SNP-Seek database (v3) software for viability for PH Rice GDB
- Genotyping Data Preparation and Transformation
- Variety and phenotyping data preparation and transformation
📜 Deliverables
Deliverable 1: Current SNP-Seek DB Evaluation, Gap Analysis, and Development Roadmap
1. Review SNP-Seek v3 Features
- Identify key features and functionalities required by PH Rice GDB
- Confirm feature alignment with user needs (from project team discussions)
2. Set Up Software Development Version Control
- Create a project directory structure
- Organize configurations, components, utilities, and assets
- Identify and design initial API structure, use cases, and endpoints required by PH Rice GDB
- Identify additional use cases not in the current design but important to PH Rice GDB end users
🎯 Tangible Results
- Feature review document (includes API/Use Case Design Document)
- Code Repository (Bitbucket or GitHub) with the initial project structure
Genotyping data for subset 3KRG accessions from PH
- Clean, normalize, and validate data (VCFs)
- Transform data into a suitable intermediate format (prior to HDF5 creation)
- Generate HDF5 files for efficient storage and retrieval of structured data
- Prepare and curate passport and agro-morphological data for 1k1 TRVs
- Transform variety data for application loading
🎯 Tangible Results
- Scripts for data validation and transformation, uploaded to repository
- Transformed data (genotype) in required formats
- HDF5 files with structured data
- Raw Data - Sample Raw Data and outputs