zhengxwen/gdsfmt: R Interface to CoreArray Genomic Data Structure (GDS) Files
Version 1.17.4

This package provides a high-level R interface to CoreArray Genomic Data Structure (GDS) data files, which are portable across platforms with hierarchical structure to store multiple scalable array-oriented data sets with metadata information. It is suited for large-scale datasets, especially for data which are much larger than the available random-access memory. The gdsfmt package offers the efficient operations specifically designed for integers of less than 8 bits, since a diploid genotype, like single-nucleotide polymorphism (SNP), usually occupies fewer bits than a byte. Data compression and decompression are available with relatively efficient random access. It is also allowed to read a GDS file in parallel with multiple R processes supported by the package parallel.

Getting started

Package details

AuthorXiuwen Zheng [aut, cre], Stephanie Gogarten [ctb], Jean-loup Gailly and Mark Adler [ctb] (for the included zlib sources), Yann Collet [ctb] (for the included LZ4 sources), xz contributors (for the included liblzma sources)
Bioconductor views DataImport Infrastructure Software
MaintainerXiuwen Zheng <[email protected]>
LicenseLGPL-3
Version1.17.4
URL http://corearray.sourceforge.net/ http://github.com/zhengxwen/gdsfmt
Package repositoryView on GitHub
Installation Install the latest version of this package by entering the following in R:
install.packages("devtools")
library(devtools)
install_github("zhengxwen/gdsfmt")
zhengxwen/gdsfmt documentation built on Sept. 2, 2018, 11:20 p.m.