“ASR:2015-07-06”版本间的差异

来自cslt Wiki
跳转至: 导航搜索
(以“==Speech Processing == === AM development === ==== Environment ==== ==== RNN AM==== *morpheme RNN --zhiyuan ==== Mic-Array ==== * hold * compute EER with kaldi...”为内容创建页面)
 
Text Processing
第50行: 第50行:
 
* (hold)
 
* (hold)
  
===Order representation ===
+
====Order representation ====
 
* Nested Dropout
 
* Nested Dropout
 
* modify the objective function(hold)
 
* modify the objective function(hold)
===Balance Representation===
+
====Balance Representation====
 
* Find error signal
 
* Find error signal
  
===Recommendation===
+
====Recommendation====
 
* Reproduce baseline.
 
* Reproduce baseline.
  
===DSSM based QA===
+
====DSSM based QA====
 
* Reproduce baseline.
 
* Reproduce baseline.

2015年7月6日 (一) 01:34的版本

Speech Processing

AM development

Environment

RNN AM

  • morpheme RNN --zhiyuan


Mic-Array

  • hold
  • compute EER with kaldi

====Data selection unsupervised learning

  • train using aurora4 --zhiyong
  • train using wsj --xuewei

RNN-DAE(Deep based Auto-Encode-RNN)

  • hold
  • deliver to mengyuan

Speaker ID

  • DNN-based sid --Lantian


Ivector&Dvector based ASR

  • hold --Tian Lan
  • Cluster the speakers to speaker-classes, then using the distance or the posterior-probability as the metric
  • dark-konowlege using i-vector
  • train on wsj(testbase dev93+evl92)
  • --hold

Dark knowledge

  • test random last output layer when train MPE --zhiyuan


language vector

  • hold

Text Processing

RNN LM

  • character-lm rnn(hold)
  • lstm+rnn
  • check the lstm-rnnlm code about how to Initialize and update learning rate.(hold)

Neural Based Document Classification

  • (hold)

Order representation

  • Nested Dropout
  • modify the objective function(hold)

Balance Representation

  • Find error signal

Recommendation

  • Reproduce baseline.

DSSM based QA

  • Reproduce baseline.