icefall

Author	SHA1	Message	Date
Daniel Povey	5f2c0a09b7	Convert swish nonlinearities to ReLU	2022-03-05 16:28:24 +08:00
Daniel Povey	0cd14ae739	Fix exp dir	2022-03-05 12:17:09 +08:00
Daniel Povey	65b09dd5f2	Double the threshold in brelu; slightly increase max_factor.	2022-03-05 00:07:14 +08:00
Daniel Povey	74f2b163de	Merge diagnostics improvement	2022-03-04 23:15:47 +08:00
Daniel Povey	6252282fd0	Add deriv-balancing code	2022-03-04 20:19:11 +08:00
Daniel Povey	eb3ed54202	Reduce scale from 50 to 20	2022-03-04 15:56:45 +08:00
Daniel Povey	9cc5999829	Fix duplicate Swish; replace norm+swish with swish+exp-scale in convolution module	2022-03-04 15:50:51 +08:00
yaozengwei	ad62981765	Add diagnostics (#230 ) * Adding diagnostics code... * Move diagnostics code from local dir to the shared icefall dir * Remove the diagnostics code in the local dir * Update docs of arguments, and remove stats_types() function in TensorDiagnosticOptions object. * Update docs of arguments. * Add copyright information. * Corrected the time in copyright information. Co-authored-by: Daniel Povey <dpovey@gmail.com>	2022-03-04 15:38:23 +08:00
Daniel Povey	7e88999641	Increase scale from 20 to 50.	2022-03-04 14:31:29 +08:00
Daniel Povey	3207bd98a9	Increase scale on Scale from 4 to 20	2022-03-04 13:16:40 +08:00
Daniel Povey	503f8d521c	Fix bug in diagnostics	2022-03-04 13:08:56 +08:00
Daniel Povey	3d9ddc2016	Fix backprop bug	2022-03-04 12:29:44 +08:00
Fangjun Kuang	2f0fbf430c	Remove duplicate files. (#236 )	2022-03-04 11:56:31 +08:00
Daniel Povey	cd216f50b6	Add import	2022-03-04 11:03:01 +08:00
Daniel Povey	bc6c720e25	Combine ExpScale and swish for memory reduction	2022-03-04 10:52:05 +08:00
Daniel Povey	23b3aa233c	Double learning rate of exp-scale units	2022-03-04 00:42:37 +08:00
Daniel Povey	5c177fc52b	pelu_base->expscale, add 2xExpScale in subsampling, and in feedforward units.	2022-03-03 23:52:03 +08:00
Fangjun Kuang	3ec219dfa0	Add stateless transducer tutorial. (#235 ) * WIP: Add stateless transducer tutorial. * Add more doc. * Minor fixes.	2022-03-03 22:33:47 +08:00
Daniel Povey	3fb559d2f0	Add baseline for the PeLU expt, keeping only the small normalization-related changes.	2022-03-02 18:27:08 +08:00
Fangjun Kuang	1ff6196c44	Fix joiner (#234 ) * Add tests for Joiner * Remove duplicate files.	2022-03-02 16:41:14 +08:00
Daniel Povey	9ed7d55a84	Small bug fixes/imports	2022-03-02 16:34:55 +08:00
Daniel Povey	9d1b4ae046	Add pelu to this good-performing setup..	2022-03-02 16:33:27 +08:00
Fangjun Kuang	50d2281524	Add modified transducer loss for AIShell dataset (#219 ) * Add modified transducer for aishell. * Minor fixes. * Add extra data in transducer training. The extra data is from http://www.openslr.org/62/ * Update export.py and pretrained.py * Update CI to install pretrained models with aishell. * Update results. * Update results. * Update README. * Use symlinks to avoid copies.	2022-03-02 16:02:38 +08:00
Fangjun Kuang	05cb297858	Update result for full libri + GigaSpeech using transducer_stateless. (#231 )	2022-03-01 17:01:46 +08:00
Fangjun Kuang	72f838dee1	Update results for transducer_stateless after training for more epochs. (#207 )	2022-03-01 16:35:02 +08:00
Daniel Povey	2ff520c800	Improvements to diagnostics (RE those with 1 dim	2022-02-28 12:22:27 +08:00
Daniel Povey	c1063def95	First version of rand-combine iterated-training-like idea.	2022-02-27 17:34:58 +08:00
Daniel Povey	63d8d935d4	Refactor/simplify ConformerEncoder	2022-02-27 13:56:15 +08:00
Daniel Povey	581786a6d3	Adding diagnostics code...	2022-02-27 13:44:43 +08:00
PF Luo	ac7c2d84bc	minor fix for aishell recipe (#223 ) * just remove unnecessary torch.sum * minor fixs for aishell	2022-02-23 08:33:20 +08:00
Fangjun Kuang	2332ba312d	Begin to use multiple datasets in training (#213 ) * Begin to use multiple datasets. * Finish preparing training datasets. * Minor fixes * Copy files. * Finish training code. * Display losses for gigaspeech and librispeech separately. * Fix decode.py * Make the probability to select a batch from GigaSpeech configurable. * Update results. * Minor fixes.	2022-02-21 15:27:27 +08:00
Fangjun Kuang	1c35ae1dba	Reset seed at the beginning of each epoch. (#221 ) * Reset seed at the beginning of each epoch. * Use a different seed for each epoch.	2022-02-21 15:16:39 +08:00
Fangjun Kuang	cbf8c18ebd	Minor fixes for aishell (#218 ) * Minor fixes to aishell. * Minor fixes.	2022-02-19 22:28:19 +08:00
PF Luo	277cc3f9bf	update aishell-1 recipe with k2.rnnt_loss (#215 ) * update aishell-1 recipe with k2.rnnt_loss * fix flak8 style * typo * add pretrained model link to result.md	2022-02-19 15:56:39 +08:00
Duo Ma	827b9df51a	Updated Aishell-1 transducer-stateless result (#217 ) * Update RESULTS.md * Update RESULTS.md	2022-02-19 15:56:04 +08:00
Wei Kang	b702281e90	Use k2 pruned transducer loss to train conformer-transducer model (#194 ) * Using k2 pruned version transducer loss to train model * Fix style * Minor fixes	2022-02-17 13:33:54 +08:00
Daniel Povey	2af1b3af98	Remove ReLU in attention	2022-02-14 19:39:19 +08:00
Daniel Povey	d187ad8b73	Change max_frames from 0.2 to 0.15	2022-02-11 16:24:17 +08:00
Daniel Povey	4cd2c02fff	Fix num_time_masks code; revert 0.8 to 0.9	2022-02-10 15:53:11 +08:00
Daniel Povey	c170c53006	Change p=0.9 to p=0.8 in SpecAug	2022-02-10 14:59:14 +08:00
Daniel Povey	8aa50df4f0	Change p=0.5->0.9, mask_fraction 0.3->0.2	2022-02-09 22:52:53 +08:00
Wang, Guanbo	70a3c56a18	Fix librispeech train.py (#211 ) * fix librispeech train.py * remove note	2022-02-09 16:42:28 +08:00
Daniel Povey	dd19a6a2b1	Fix to num_feature_masks bug I introduced; reduce max_frames_mask_fraction 0.4->0.3	2022-02-09 12:02:19 +08:00
Daniel Povey	bd36216e8c	Use much more aggressive SpecAug setup	2022-02-08 21:55:20 +08:00
Daniel Povey	beaf5bfbab	Merge specaug change from Mingshuang.	2022-02-08 19:42:23 +08:00
Daniel Povey	395065eb11	Merge branch 'spec-augment-change' of https://github.com/luomingshuang/icefall into attention_relu_specaug	2022-02-08 19:40:33 +08:00
Mingshuang Luo	3323cabf46	Experiments based on SpecAugment change	2022-02-08 14:25:31 +08:00
Fangjun Kuang	27fa5f05d3	Update git SHA-1 in RESULTS.md for transducer_stateless. (#202 )	2022-02-07 18:45:45 +08:00
Fangjun Kuang	a8150021e0	Use modified transducer loss in training. (#179 ) * Use modified transducer loss in training. * Minor fix. * Add modified beam search. * Add modified beam search. * Minor fixes. * Fix typo. * Update RESULTS. * Fix a typo. * Minor fixes.	2022-02-07 18:37:36 +08:00
Daniel Povey	a859dcb205	Remove learnable offset, use relu instead.	2022-02-07 12:14:48 +08:00

... 4 5 6 7 8 ...

413 Commits