fast-bwa

Commit Graph

Author	SHA1	Message	Date
zzh	b50360bb48	整合了ert seeding，结果正确，完善了shm	2024-08-06 03:08:10 +08:00
zzh	364ee9756e	添加了共享内存支持，成功构建ert索引	2024-08-03 17:26:32 +08:00
zzh	8c2a90e3e2	修改了双线程解压读写fastq.gz的bug，修改了一些测试代码	2024-06-24 18:11:54 +08:00
zzh	d41c038616	清理一些fprintf	2024-04-06 16:08:43 +08:00
zzh	dd7db7beb6	修改一下索引文件后缀名称	2024-04-02 07:50:22 +08:00
zzh	20e072f6af	所有代码都合并了，还差一点建立索引的时候，一起都建立了	2024-04-02 07:42:37 +08:00
zzh	1e3965cb7d	添加了可同时读写的pipeline，优化了时间统计	2024-03-24 04:40:09 +08:00
zzh	7d085962a2	开始改成sbwa那种batch模式	2024-03-07 18:23:21 +08:00
zzh	6e1dd08fb6	将seed和extend部分修改成了batch模式，好像没啥效果	2024-02-23 01:09:08 +08:00
zzh	4bd0fd4f91	做了一些代码清理，目前结果应该是完全一致的	2024-02-22 01:26:57 +08:00
zzh	17618ee5f2	解决了sa的bug，现在结果和原版一模一样	2024-02-21 15:21:56 +08:00
zzh	fc2e0d9b0b	实现了seed过程的所有加速想法，seed部分实现了3倍左右加速比	2024-02-20 01:12:02 +08:00
zzh	9d45fd02fb	添加了bit过滤，解决了一些bug，现在seed1和seed2都没问题了	2024-02-16 00:18:14 +08:00
zzh	d41b8da061	kmer长度变为14，结果正确	2024-02-12 20:54:57 +08:00
zzh	463f7da138	将smem1函数用fmt结构实现了，结果基本正确	2024-02-07 22:08:51 +08:00
zzh	bf678f4dae	实现了用33bit表示sa，间隔为4，释放内存的时候会崩溃	2023-12-27 10:42:12 +08:00
Nils Homer	56026158d8	Add the header line to the output SAM In particular, this defines the output SAM to be unsorted BUT also query grouped. The latter is very important to explicitly define so downstream tools that don't make assumptions know that reads from the same template are grouped.	2021-12-14 08:02:05 -07:00
Heng Li	d422bdbed9	debug flag to measure memory	2021-02-22 23:26:03 -05:00
Heng Li	02a9add042	added MIT license to some non-GPL source files	2020-07-01 23:02:01 -04:00
Heng Li	1eee77ad93	Merge pull request #84 from jblachly/readgroupfix Forbid literal TAB control characters in @RG line	2017-07-30 18:56:39 -04:00
John Marshall	690649872b	Copy the whole kstring_t even if it contains NULs FASTQ files containing NULs are invalid but should not cause bwa to crash, as it does if the quality line contains a NUL. Fixes #122.	2017-06-30 12:46:56 +01:00
John Marshall	ab3a92bc73	Prevent Clang warnings on abs() and fabs() calls In the bwa.c and bwase.c calls, rlen is an int64_t returned from bns_get_seq() and is the number of reference bases covered by the alignment; l_query/len is an int and the query length of the alignment; and the result is an int given to an int parameter of ksw_global[2](). As even the result is int and as rlen is effectively bounded by the maximum length of a reference sequence, we maintain the status quo in this code and simply cast rlen to int to silence Clang's "use llabs()" (llabs() would not be a great answer given an int64_t anyway). The bwtsw2_pair.c call needs to remain fabs() so both divisions are done in floating point; cast to double to prevent Clang suggesting changing the call to integer abs().	2017-06-26 10:45:13 +01:00
James Blachly	8cc02badd9	Forbid literal TAB control characters in @RG line	2016-08-09 17:05:51 -04:00
Heng Li	7ec3261877	r1134: use AH:* instead of AH:Y	2016-05-03 11:28:58 -04:00
Heng Li	3c038250f9	r1130: changed "ah" to "AH"	2016-04-28 15:39:18 -04:00
Heng Li	c561759222	r1027: segfault caused by the last commit	2014-12-12 16:56:54 -05:00
Heng Li	925ddfb697	r1025: accept file with -H; allow to replace @SQ	2014-12-11 10:38:36 -05:00
Heng Li	b5f6ed3020	r1005: insert arbitrary header lines	2014-11-19 10:59:05 -05:00
Heng Li	80e4ecfa79	r998: smart pairing; allow mixture of SE/PE reads	2014-11-18 14:30:22 -05:00
Heng Li	a06646493b	r915: fixed broken example.c	2014-10-17 16:17:28 -04:00
Heng Li	e318d8e7e5	r905: lower peak RAM for "shm -f"	2014-10-16 11:22:09 -04:00
Heng Li	bfd5e1840f	shm works on small files, but not large ones I don't know why. SHMMAX, SHMALL and SHMMNI are large enough.	2014-10-15 15:44:06 -04:00
Heng Li	6a0952948d	shared memory	2014-10-15 14:44:08 -04:00
Heng Li	c5e859b49f	r898: read the index into a single memory block Prepare for shared memory. Not used now.	2014-10-15 12:27:45 -04:00
Heng Li	71277f0fea	r896: more flexible ALT reading	2014-10-14 23:37:24 -04:00
Heng Li	7954e77a1b	r741: fixed segfault in rare cases	2014-05-01 11:13:05 -04:00
Heng Li	b93fca2b2e	r723: merge adjacent hits	2014-04-16 16:38:50 -04:00
Heng Li	8638cfadc8	dev-472: get rid of bwa_fix_xref() This function causes all kinds of problems when the reference genome consists of many short reads/contigs/chromsomes. Some of the problems are nearly unfixable at the point where bwa_fix_xref() gets called. This commit attempts to fix the problem at the root. It disallows chains spanning multiple contigs and never retrieves sequences bridging two adjacent contigs. Thus all the chaining, extension, SW and global alignments are confined to on contig only. This commit brings many changes. I have tested it on a couple examples including Peter Field's PacBio example. It works well so far.	2014-04-10 20:54:27 -04:00
Heng Li	ccbbe48c4f	dev-470: don't stop on bwa_fix_xref2() failures Peter Field has sent me an example caused by an alignment bridging three adjacent chromosomes/contigs. Bwa-mem always aligns the query to the contig covering the middle point of the alignment. In this example, it chooses the middle contig, which should not be aligned. This leads to weird things failing bwa_fix_xref2(), which cannot be fixed unless we build the contig boundaries into the FM-index. In the old code, bwa-mem halts when bwa_fix_xref2() fails. With this commit, bwa-mem will give a warning instead of halting.	2014-04-10 11:43:17 -04:00
Heng Li	9ce50a4e5e	dev-450: support diff ins/del penalties. NO TEST!!	2014-03-28 14:54:06 -04:00
Heng Li	7d63e76245	r444: more debugging output in CIGAR generation Also found a potential issue which should not affect accuracy but may hurt speed. Will investigate later.	2014-03-16 23:25:04 -04:00
Heng Li	e879817373	r440: a condition not work due to a typo	2014-02-20 13:06:40 -05:00
Heng Li	17fb85a227	r438: still an issue in MD It occurs when the global alignment disagrees with the local alignment.	2014-02-19 11:31:54 -05:00
Heng Li	bdd14d2946	r436: fix rare MD/NM-CIGAR inconsistencies	2014-02-19 10:08:43 -05:00
Heng Li	4adc34eccb	r435: bugfix - base not complemented on the rev	2014-02-18 10:32:24 -05:00
Heng Li	7c50bad567	Release bwa-0.7.6a-r433	2014-01-31 12:58:21 -05:00
Heng Li	f524c7d3d8	r431: added the MD tag to bwa-mem	2014-01-29 12:05:11 -05:00
Heng Li	ff6faf811a	r419: print the @PG line	2013-11-19 11:08:45 -05:00
Heng Li	9735d7a31a	conform to the latest (unpublished) SAM spec for chimeric alignments	2013-05-22 19:45:16 -04:00
Heng Li	9a6abe51b6	r391: better method to resolve xref alignment The old method does not work when the alignment bridges three chr. This may actually happen often. The new method does not work all the time, either, but should be better than the old one. It is also simpler, arguably.	2013-05-22 18:57:51 -04:00

1 2

80 Commits (b50360bb488b63a4b28290d5ba5369dbd2e45f2e)