← All conferences

ICASSP 2015 Accepted Papers

The full list of 1,198 papers accepted at ICASSP 2015 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

  1. Sampling theory for graph signals
  2. Scalable audio separation with light Kernel Additive Modelling
  3. Scalable clustering based on enhanced-SMART for large-scale FMRI datasets
  4. Scale-robust compressive camera fingerprint matching with random projections
  5. Scaling recurrent neural network language models
  6. Searching for semantic person queries using channel representations
  7. Second order statistics of bilinear forms of robust scatter estimators
  8. Secrecy rate analysis for jamming assisted relay communications systems
  9. Section-level modeling of musical audio for linking performances to scores in Turkish makam music
  10. Security information factor based low probability of identification in distributed multiple-radar system
  11. Segmental acoustic indexing for zero resource keyword search
  12. Seismic feature extraction using steiner tree methods
  13. Selective hole-filling for depth-image based rendering
  14. Selective video encryption using chaotic system in the SHVC extension
  15. Self-calibration in visual sensor networks equipped with RGB-D cameras
  16. Semi supervised deep kernel design for image annotation
  17. Semi-asynchronous routing for large scale hierarchical networks
  18. Semi-supervised multi-sensor classification via consensus-based Multi-View Maximum Entropy Discrimination
  19. Semi-supervised training in low-resource ASR and KWS
  20. Sensor selection with correlated measurements for target tracking in wireless sensor networks
  21. Separating background and foreground optical flow fields by low-rank and sparse regularization
  22. Sequence-discriminative training of recurrent neural networks
  23. Sequential energy detection for touch input detection
  24. Serial and interleaved architectures for computing real FFT
  25. Session negotiation and media adaptation of EVS in Voice over LTE
  26. Shape peeling for improved image Skeleton stability
  27. Signal processing considerations for passive radar with a single receiver
  28. Signal processing on graphs: Estimating the structure of a graph
  29. Similarity induced group sparsity for non-negative matrix factorisation
  30. Singing voice analysis and editing based on mutually dependent F0 estimation and source separation
  31. Singing voice detection with deep recurrent neural networks
  32. Single carrier with multi-channel time-frequency domain equalization for underwater acoustic communications
  33. Single channel speech enhancement in the modulation domain: New insights in the modulation channel selection framework
  34. Single image haze removal via a simplified dark channel
  35. Single stream parallelization of generalized LSTM-like RNNs on a GPU
  36. Single underwater image descattering and color correction
  37. Single-channel blind estimation of reverberation parameters
  38. Single-channel speech enhancement in a transient noise environment by exploiting speech harmonicity
  39. Small target detection using an optimization-based filter
  40. Small-footprint high-performance deep neural network-based speech recognition using split-VQ
  41. Smelly parallel MCMC chains
  42. Softsad: Integrated frame-based speech confidence for speaker recognition
  43. Sound event detection in real life recordings using coupled matrix factorization of spectral representations and class activity annotations
  44. Source counting in speech mixtures by nonparametric Bayesian estimation of an infinite Gaussian mixture model
  45. Source separation with scattering Non-Negative Matrix Factorization
  46. Source-specific informative prior for i-vector extraction
  47. Space-delay adaptive processing for MIMO RF indoor motion mapping
  48. Sparse HMM-based speech enhancement method for stationary and non-stationary noise environments
  49. Sparse and cross-term free time-frequency distribution based on Hermite functions
  50. Sparse and low rank decomposition using l0 penalty
  51. Sparse chroma estimation for harmonic audio
  52. Sparse models for determining arterial dynamics
  53. Sparse null space basis pursuit and analysis dictionary learning for high-dimensional data analysis
  54. Sparse partial derivatives and reconstruction from partial Fourier data
  55. Sparse representation for frequency warping based voice conversion
  56. Sparse sensing for distributed gaussian detection
  57. Sparse signal recovery in the presence of colored noise and rank-deficient noise covariance matrix: An SBL approach
  58. Sparse symbol detection by a greedy tree search
  59. Sparsity aware minimum error entropy algorithms
  60. Sparsity pattern recovery using FRI methods
  61. Spatial diffuseness features for DNN-based speech recognition in noisy and reverberant environments
  62. Spatio-temporal rich model for motion vector steganalysis
  63. Speaker adaptive training for deep neural networks embedding linear transformation networks
  64. Speaker and noise independent online single-channel speech enhancement
  65. Speaker change detection and speaker diarization using spatial information
  66. Speaker change point detection using deep neural nets
  67. Speaker verification with the mixture of Gaussian factor analysis based representation
  68. Spectral conversion using deep neural networks trained with multi-source speakers
  69. Spectral envelope reconstruction via IGF for audio transform coding
  70. Spectral mask estimation using deep neural networks for inter-sensor data ratio model based robust DOA estimation
  71. Spectral properties of neuronal pulse interval modulation
  72. Spectrum cartography using quantized observations
  73. Spectrum scanning when the intruder might have knowledge about the scanner's capabilities
  74. Spectrum sharing between matrix completion based MIMO radars and a MIMO communication system
  75. Speech Separation based on signal-noise-dependent deep neural networks for robust speech recognition
  76. Speech acoustic modeling from raw multichannel waveforms
  77. Speech dereverberation using a learned speech model
  78. Speech emotion recognition with acoustic and lexical features
  79. Speech recognition with prediction-adaptation-correction recurrent neural networks
  80. Speech reinforcement in noisy reverberant conditions under an approximation of the short-time SII
  81. Speech-codebook based soft Voice Activity Detection
  82. Speech-laughs: An HMM-based approach for amused speech synthesis
  83. Spherical harmonic transform for minimum dimensionality regular grid sampling on the sphere
  84. Spikes from compound action potentials in simulated microelectrode recordings
  85. Stability analysis of the FBANC system having delay error in the estimated secondary path model
  86. Stability and continuity of centrality measures in weighted graphs
  87. Stabilization techniques for high resolution ultrasound imaging using beamspace Capon method
  88. Standardization of the new 3GPP EVS codec
  89. Statistical modeling of binaural signal and its application to binaural source separation
  90. Statistical-mechanical analysis of the FXLMS algorithm with actual primary path
  91. Structural segmentation of Hindustani concert audio with posterior features
  92. Structure discovery of deep neural network based on evolutionary algorithms
  93. Structured Bayesian compressive sensing exploiting spatial location dependence
  94. Structured sparse signal models and decomposition algorithm for super-resolution in sound field recording and reproduction
  95. Subjective quality evaluation of the 3GPP EVS codec
  96. Submodular data selection with acoustic and phonetic features for automatic speech recognition
  97. Subspace leakage analysis of sample data covariance matrix
  98. Subspace learning using consensus on the grassmannian manifold
  99. Subspace projection matrix completion on Grassmann manifold
  100. Subspace-based phase noise estimation in OFDM receivers
  101. Sum rate maximization model of non-regenerative multi-stream multi-pair multi-relay network
  102. Super-resolution acoustic imaging using sparse recovery with spatial priming
  103. Super-resolution in Phase Space
  104. Super-resolution ultrawideband ultrasound imaging using focused frequency time reversal music
  105. Super-wideband bandwidth extension for speech in the 3GPP EVS codec
  106. Supervised domain adaptation for emotion recognition from speech
  107. Supervised hierarchical segmentation for bird song recording
  108. Supervised sparse coding with local geometrical constraints
  109. Support knowledge-aided sparse Bayesian learning for compressed sensing
  110. Switching dual kernels for separable edge-preserving filtering
  111. Switching to and combining offline-adapted cluster acoustic models based on unsupervised segment classification
  112. Synchronization rules for HMM-based audio-visual laughter synthesis
  113. System architectures and digital signal processing algorithms for enhancing the output audio quality of stereo FM broadcast receivers
  114. Telephony text-prompted speaker verification using i-vector representation
  115. Temporal Tile Shaping for spectral gap filling in audio transform coding in EVS
  116. Temporal entropy-based texturedness indicator for audio signals
  117. Tensor object classification via multilinear discriminant analysis network
  118. The AMG1608 dataset for music emotion recognition
  119. The THUEE system for the openKWS14 keyword search evaluation
  120. The effect of neural networks in statistical parametric speech synthesis
  121. The efficiency of view synthesis prediction for 3D video coding: A spectral domain analysis
  122. The proportional mean decomposition: A bridge between the Gaussian and bernoulli ensembles
  123. The role of glottal source parameters for high-quality transformation of perceptual age
  124. The segregation of spatialised speech in interference by optimal mapping of diverse cues
  125. The shared dirichlet priors for bayesian language modeling
  126. The widely linear quaternion recursive total least squares
  127. Tikhonov-Galerkin stochastic system identification in SO(3)
  128. Time-frequency image descriptors-based features for EEG epileptic seizure activities detection and classification
  129. Time-reversal space-time codes in asynchronous two-way double-antenna relay networks
  130. Time-switching based SWPIT for network-coded two-way relay transmission with data rate fairness
  131. Time-varying vector Poisson processes with coincidences
  132. Token-level interpolation for class-based language models
  133. Tokenizing fundamental frequency variation for Mandarin tone error detection
  134. Tonal complexity features for style classification of classical music
  135. Topological interference management for two cell interference broadcast channels with alternating connectivity
  136. Total Jensen divergences: Definition, properties and clustering
  137. Total generalized variation for graph signals
  138. Towards machines that know when they do not know: Summary of work done at 2014 Frederick Jelinek Memorial Workshop
  139. Tracking changes in functional connectivity of brain networks from resting-state fMRI using particle filters
  140. Transient interference suppression via structured low-rank matrix decomposition
  141. Transmission distortion modeling for view synthesis prediction based 3-D video streaming
  142. Transmit code design for extended target detection in the presence of clutter
  143. Transmitting informative components of fisher codes for mobile visual search
  144. Trinicon-BSS system incorporating robust dual beamformers for noise reduction
  145. Twice-universal piecewise linear regression via infinite depth context trees
  146. Two-stage speech/music classifier with decision smoothing and sharpening in the EVS codec
  147. Tyler's estimator performance analysis
  148. Under-sampled functional MRI using low-rank plus sparse matrix decomposition
  149. Unicode-based graphemic systems for limited resource languages
  150. Unidirectional long short-term memory recurrent neural network with recurrent output layer for low-latency speech synthesis
  151. Unit circle MVDR beamformer
  152. Universal lower bounds on sampling rates for covariance estimation
  153. Universal outlier hypothesis testing: Application to anomaly detection
  154. Unnormalized exponential and neural network language models
  155. Unscented Transformation based array interpolation
  156. Unsupervised adaptation of a denoising autoencoder by Bayesian Feature Enhancement for reverberant asr under mismatch conditions
  157. Unsupervised data selection and word-morph mixed language model for tamil low-resource keyword search
  158. Unsupervised detection of malware in persistent web traffic
  159. Unsupervised detrending technique using sparse dictionary learning for fMRI preprocessing and analysis
  160. Unsupervised feature learning for urban sound classification
  161. Unsupervised learning of acoustic features via deep canonical correlation analysis
  162. Unsupervised neural network based feature extraction using weak top-down constraints
  163. Unsupervised speaker adaptation of deep neural network based on the combination of speaker codes and singular value decomposition for speech recognition
  164. Unusual event detection in crowded scenes by trajectory analysis
  165. Unveiling the tree: A convex framework for sparse problems
  166. Utilizing spectro-temporal correlations for an improved speech presence probability based noise power estimation
  167. Variational Bayes learning of multiscale graphical models
  168. Variational Bayes state space model for acoustic echo reduction and dereverberation
  169. Variational EM for clustering interaural phase cues in MESSL for blind source separation of speech
  170. Variational inference cooperative network localization with narrowband radios
  171. View synthesis optimization based on texture smoothness for 3D-HEVC
  172. Visual and acoustic identification of bird species
  173. Visual tracking using learned color features
  174. Visualization of sound field by means of Schlieren method with spatio-temporal filtering
  175. Vital signs from inside a helmet: A multichannel face-lead study
  176. Vocaine the vocoder and applications in speech synthesis
  177. Vocal activity informed singing voice separation with the iKala dataset
  178. Vocal responses to frequency modulated composite sinewaves via auditory and vibrotactile pathways
  179. Voice activity detection using subband noncircularity
  180. Voice conversion using deep Bidirectional Long Short-Term Memory based Recurrent Neural Networks
  181. Voice quality: Not only about "you" but also about "your interlocutor"
  182. Voltage sags estimation in three-phase systems using Unconditional Maximum Likelihood estimation
  183. WFST-based structural classification integrating dnn acoustic features and RNN language features for speech recognition
  184. Wave atom based Compressive Sensing and adaptive beamforming in ultrasound imaging
  185. Wavelet-based compressed spectrum sensing for cognitive radio wireless networks
  186. Weak interference direction of arrival estimation in the GPS L1 frequency band
  187. Weight estimation in hypergraph learning
  188. Weighted covariance matching based square root LASSO
  189. Weighted one-norm minimization with inaccurate support estimates: Sharp analysis via the null-space property
  190. Weighted pairwise Gaussian likelihood regression for depression score prediction
  191. Weighted training for speech under Lombard Effect for speaker recognition
  192. Wideband waveform design for robust target detection
  193. Wind noise short term power spectrum estimation using pitch adaptive inverse binary masks
  194. Wireless information and power transfer in MIMO channels under Rician fading
  195. Word embedding for recurrent neural network based TTS synthesis
  196. Word-semantic lattices for spoken language understanding
  197. eTutor: Online learning for personalized education
  198. ℓ1-constrained MVDR-based selection of nonidentical directivities in microphone array

Looking for submission deadlines instead? See the conference deadline calendar.