Skip to content
NBDC Human Database

No datasets in the cart.

Due to system maintenance, the application system, application review by the Data Access Committee will be unavailable during the following period.
Schedule: October 5th (Mon), 2026, 9:00 - October 7th (Wed), 2026, 15:00 (JST)
We apologize for any inconvenience this may cause and appreciate your understanding.

We are currently receiving a large number of applications for data submission, and the review process is taking longer than usual.We sincerely apologize for the delay and kindly ask for your understanding. When submitting an application, we would greatly appreciate it if you could allow sufficient time for the processing.

Following a change to our organizational structure effective April 1, 2026, this division has been renamed from the "Database Center for Life Science, Joint Support-Center for Data Science Research" to the "Database Division for Life Science (DBCLS), BioData Science Initiative (BSI), National Institute of Genetics (NIG)". Where the former name still appears in the guidelines, please read it as the new name.

Dataset ID

NHA000190

Type of data
GWAS for gut microbiome
GWAS for plasma metabolite
GWAS for KEGG Gene Ortholog and KEGG Pathway
Access criteria
Unrestricted-access
Total data volume
320 GB
File formats
  • TSV
  • DOCX
Research
hum0197
Date published
2023-10-02
Date modified
2023-10-02
Secondary ID
hum0197.v18.gwas.v1

Unrestricted-access files linked to this dataset

Per page
50

101–150 / 732

FileLabelSizeCopy URL
metabo_C_0068_QCed_sumstats.tsv.gzUrocanic acid318 MB
metabo_C_0070_QCed_sumstats.tsv.gz1-Methyl-4-imidazoleacetic acid319 MB
metabo_C_0072_QCed_sumstats.tsv.gzXC0029319 MB
metabo_C_0073_QCed_sumstats.tsv.gzStachydrine318 MB
metabo_C_0074_QCed_sumstats.tsv.gzγ-Butyrobetaine318 MB
metabo_C_0076_QCed_sumstats.tsv.gzGln318 MB
metabo_C_0077_QCed_sumstats.tsv.gzLys318 MB
metabo_C_0078_QCed_sumstats.tsv.gzN-Acetylserine318 MB
metabo_C_0079_QCed_sumstats.tsv.gzGlu318 MB
metabo_C_0080_QCed_sumstats.tsv.gzMet319 MB
metabo_C_0081_QCed_sumstats.tsv.gzTriethanolamine318 MB
metabo_C_0086_QCed_sumstats.tsv.gzHis318 MB
metabo_C_0087_QCed_sumstats.tsv.gzImidazolelactic acid318 MB
metabo_C_0091_QCed_sumstats.tsv.gzN6-Methyllysine319 MB
metabo_C_0092_QCed_sumstats.tsv.gzO-Acetylhomoserine;2-Aminoadipic acid318 MB
metabo_C_0093_QCed_sumstats.tsv.gzCarnitine318 MB
metabo_C_0094_QCed_sumstats.tsv.gz5-Hydroxylysine318 MB
metabo_C_0095_QCed_sumstats.tsv.gzS-Methylmethionine318 MB
metabo_C_0096_QCed_sumstats.tsv.gzMethionine sulfoxide318 MB
metabo_C_0097_QCed_sumstats.tsv.gzPhe318 MB
metabo_C_0100_QCed_sumstats.tsv.gz1-Methylhistidine;3-Methylhistidine319 MB
metabo_C_0103_QCed_sumstats.tsv.gzArg318 MB
metabo_C_0105_QCed_sumstats.tsv.gzGuanidinosuccinic acid318 MB
metabo_C_0106_QCed_sumstats.tsv.gzCitrulline318 MB
metabo_C_0107_QCed_sumstats.tsv.gzSerotonin318 MB
metabo_C_0108_QCed_sumstats.tsv.gzAlliin318 MB
metabo_C_0109_QCed_sumstats.tsv.gzGluconolactone318 MB
metabo_C_0111_QCed_sumstats.tsv.gzMannosamine318 MB
metabo_C_0112_QCed_sumstats.tsv.gzGalactosamine;Glucosamine319 MB
metabo_C_0114_QCed_sumstats.tsv.gzTheobromine318 MB
metabo_C_0115_QCed_sumstats.tsv.gzParaxanthine318 MB
metabo_C_0116_QCed_sumstats.tsv.gzTyr319 MB
metabo_C_0117_QCed_sumstats.tsv.gzPhosphorylcholine318 MB
metabo_C_0119_QCed_sumstats.tsv.gzN6-Acetyllysine318 MB
metabo_C_0121_QCed_sumstats.tsv.gzNω-Methylarginine319 MB
metabo_C_0122_QCed_sumstats.tsv.gzHomoarginine318 MB
metabo_C_0123_QCed_sumstats.tsv.gzHomocitrulline318 MB
metabo_C_0124_QCed_sumstats.tsv.gzGly-Asp318 MB
metabo_C_0126_QCed_sumstats.tsv.gzCaffeine318 MB
metabo_C_0129_QCed_sumstats.tsv.gz11-Aminoundecanoic acid318 MB
metabo_C_0130_QCed_sumstats.tsv.gzAsymmetric dimethylarginine319 MB
metabo_C_0131_QCed_sumstats.tsv.gzSymmetric dimethylarginine318 MB
metabo_C_0133_QCed_sumstats.tsv.gzO-Acetylcarnitine318 MB
metabo_C_0134_QCed_sumstats.tsv.gzTrp318 MB
metabo_C_0136_QCed_sumstats.tsv.gzKynurenine319 MB
metabo_C_0137_QCed_sumstats.tsv.gz3-Methoxytyrosine319 MB
metabo_C_0138_QCed_sumstats.tsv.gzXC0061318 MB
metabo_C_0141_QCed_sumstats.tsv.gzXC0065318 MB
metabo_C_0142_QCed_sumstats.tsv.gzN-Acetylgalactosamine;N-Acetylmannosamine;N-Acetylglucosamine318 MB
metabo_C_0143_QCed_sumstats.tsv.gzCystathionine318 MB

101–150 / 732

Analysis method

genome wide SNPs

Materials and participants
524 Japanese individuals (423 species in the gut microbiome)
306 Japanese individuals (306 plasma metabolites)
524 Japanese individuals (KEGG Gene Ortholog and KEGG Pathway)
  • Subject count
    524 (Individual)
  • Population
    Japanese
Sample description
DNAs extracted from peripheral blood cells
  • Tissue
    Peripheral blood
  • Tumor / normal
    Normal
Experimental method
Genotyping by array
WGS
Reagent kit
Infinium Asian Screening Array Kit
KAPA Hyper Prep Kit
TruSeq DNA PCR-Free Library Prep Kit
Platform
Illumina HiSeq 2500
Illumina HiSeq 3000
Illumina HiSeq X
Illumina Infinium Asian Screening Array
Illumina NovaSeq 6000
Reference genome
GRCh37
QC and filtering
SNP array data:
Sample QC: We excluded individuals with low genotyping call rates (call rate < 98%). We included individuals of the estimated Asian ancestry using PCA.
Variant QC: We excluded variants with (1) genotyping call rate < 99%, (2) minor allele count < 5, (3) P-value for Hardy-Weinberg equilibrium < 1.0 × 10^−10, and (4) > 5% allele frequency difference compared with the imputation reference panel or the allele frequency panel of Tohoku Medical Megabank Project.
Post-imputation QC: We excluded imputed variants with Rsq < 0.7 and minor allele frequency < 1%.
WGS:
We excluded variants with genotype call rate <90%, ExcessHet > 60, Hardy-Weinberg P<1.0×10−10
After imputation with Beagle v5.1, we excluded imputed variants with minor allele frequency < 1%.
Imputation
Haplotype phasing: shapeit4
Imputation: minimac4
Analysis method
SNP array:
Genotyping: GenomeStudio
WGS:
WA-MEM v0.7.13 + GATK v3.8-0
PLINK2
Variant count
Gut microbiota/KEGG (SNP array): 7,213,470 variants
Blood metabolites (WGS): 6,840,258 variants
Processed data type
GWAS summary statistics
Phenotype data
Included