GitHub - raj347/BiGPY: Building similiarity graphs from large-scale biological sequence collections using Spark

BiGPy - Biological Similarity Graphs with PySpark.

bigpy-prepare is used to prepare input data to be used later for bigpy-sketch. It takes an input file or directory and processes the fasta files within to create several output files.

.brm - Lists all sequences headers from the input file(s) that were filtered during the processing stage.

.btxt - Lists all sequences, without headers, kept after filtering. A single line stores one text sequence and its ID.

.bmap - Lists all sequence headers kept after filtering with a sequence ID. ID is the line in the .bseq and .btxt files in which the sequence is listed.

Name		Name	Last commit message	Last commit date
Latest commit History 48 Commits
doc		doc
include		include
results/RUN1		results/RUN1
src		src
test		test
LICENSE_MIT.txt		LICENSE_MIT.txt
README.md		README.md
sketch.slurm		sketch.slurm

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

doc

doc

include

include

results/RUN1

results/RUN1

src

src

test

test

LICENSE_MIT.txt

LICENSE_MIT.txt

README.md

README.md

sketch.slurm

sketch.slurm

Repository files navigation

About

Releases

Packages

Languages

License

raj347/BiGPY

Folders and files

Latest commit

History

Repository files navigation

About

Resources

License

Stars

Watchers

Forks

Languages