ICASSP 2015 Accepted Papers
The full list of 1,198 papers accepted at ICASSP 2015 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.
- Sampling theory for graph signals
- Scalable audio separation with light Kernel Additive Modelling
- Scalable clustering based on enhanced-SMART for large-scale FMRI datasets
- Scale-robust compressive camera fingerprint matching with random projections
- Scaling recurrent neural network language models
- Searching for semantic person queries using channel representations
- Second order statistics of bilinear forms of robust scatter estimators
- Secrecy rate analysis for jamming assisted relay communications systems
- Section-level modeling of musical audio for linking performances to scores in Turkish makam music
- Security information factor based low probability of identification in distributed multiple-radar system
- Segmental acoustic indexing for zero resource keyword search
- Seismic feature extraction using steiner tree methods
- Selective hole-filling for depth-image based rendering
- Selective video encryption using chaotic system in the SHVC extension
- Self-calibration in visual sensor networks equipped with RGB-D cameras
- Semi supervised deep kernel design for image annotation
- Semi-asynchronous routing for large scale hierarchical networks
- Semi-supervised multi-sensor classification via consensus-based Multi-View Maximum Entropy Discrimination
- Semi-supervised training in low-resource ASR and KWS
- Sensor selection with correlated measurements for target tracking in wireless sensor networks
- Separating background and foreground optical flow fields by low-rank and sparse regularization
- Sequence-discriminative training of recurrent neural networks
- Sequential energy detection for touch input detection
- Serial and interleaved architectures for computing real FFT
- Session negotiation and media adaptation of EVS in Voice over LTE
- Shape peeling for improved image Skeleton stability
- Signal processing considerations for passive radar with a single receiver
- Signal processing on graphs: Estimating the structure of a graph
- Similarity induced group sparsity for non-negative matrix factorisation
- Singing voice analysis and editing based on mutually dependent F0 estimation and source separation
- Singing voice detection with deep recurrent neural networks
- Single carrier with multi-channel time-frequency domain equalization for underwater acoustic communications
- Single channel speech enhancement in the modulation domain: New insights in the modulation channel selection framework
- Single image haze removal via a simplified dark channel
- Single stream parallelization of generalized LSTM-like RNNs on a GPU
- Single underwater image descattering and color correction
- Single-channel blind estimation of reverberation parameters
- Single-channel speech enhancement in a transient noise environment by exploiting speech harmonicity
- Small target detection using an optimization-based filter
- Small-footprint high-performance deep neural network-based speech recognition using split-VQ
- Smelly parallel MCMC chains
- Softsad: Integrated frame-based speech confidence for speaker recognition
- Sound event detection in real life recordings using coupled matrix factorization of spectral representations and class activity annotations
- Source counting in speech mixtures by nonparametric Bayesian estimation of an infinite Gaussian mixture model
- Source separation with scattering Non-Negative Matrix Factorization
- Source-specific informative prior for i-vector extraction
- Space-delay adaptive processing for MIMO RF indoor motion mapping
- Sparse HMM-based speech enhancement method for stationary and non-stationary noise environments
- Sparse and cross-term free time-frequency distribution based on Hermite functions
- Sparse and low rank decomposition using l0 penalty
- Sparse chroma estimation for harmonic audio
- Sparse models for determining arterial dynamics
- Sparse null space basis pursuit and analysis dictionary learning for high-dimensional data analysis
- Sparse partial derivatives and reconstruction from partial Fourier data
- Sparse representation for frequency warping based voice conversion
- Sparse sensing for distributed gaussian detection
- Sparse signal recovery in the presence of colored noise and rank-deficient noise covariance matrix: An SBL approach
- Sparse symbol detection by a greedy tree search
- Sparsity aware minimum error entropy algorithms
- Sparsity pattern recovery using FRI methods
- Spatial diffuseness features for DNN-based speech recognition in noisy and reverberant environments
- Spatio-temporal rich model for motion vector steganalysis
- Speaker adaptive training for deep neural networks embedding linear transformation networks
- Speaker and noise independent online single-channel speech enhancement
- Speaker change detection and speaker diarization using spatial information
- Speaker change point detection using deep neural nets
- Speaker verification with the mixture of Gaussian factor analysis based representation
- Spectral conversion using deep neural networks trained with multi-source speakers
- Spectral envelope reconstruction via IGF for audio transform coding
- Spectral mask estimation using deep neural networks for inter-sensor data ratio model based robust DOA estimation
- Spectral properties of neuronal pulse interval modulation
- Spectrum cartography using quantized observations
- Spectrum scanning when the intruder might have knowledge about the scanner's capabilities
- Spectrum sharing between matrix completion based MIMO radars and a MIMO communication system
- Speech Separation based on signal-noise-dependent deep neural networks for robust speech recognition
- Speech acoustic modeling from raw multichannel waveforms
- Speech dereverberation using a learned speech model
- Speech emotion recognition with acoustic and lexical features
- Speech recognition with prediction-adaptation-correction recurrent neural networks
- Speech reinforcement in noisy reverberant conditions under an approximation of the short-time SII
- Speech-codebook based soft Voice Activity Detection
- Speech-laughs: An HMM-based approach for amused speech synthesis
- Spherical harmonic transform for minimum dimensionality regular grid sampling on the sphere
- Spikes from compound action potentials in simulated microelectrode recordings
- Stability analysis of the FBANC system having delay error in the estimated secondary path model
- Stability and continuity of centrality measures in weighted graphs
- Stabilization techniques for high resolution ultrasound imaging using beamspace Capon method
- Standardization of the new 3GPP EVS codec
- Statistical modeling of binaural signal and its application to binaural source separation
- Statistical-mechanical analysis of the FXLMS algorithm with actual primary path
- Structural segmentation of Hindustani concert audio with posterior features
- Structure discovery of deep neural network based on evolutionary algorithms
- Structured Bayesian compressive sensing exploiting spatial location dependence
- Structured sparse signal models and decomposition algorithm for super-resolution in sound field recording and reproduction
- Subjective quality evaluation of the 3GPP EVS codec
- Submodular data selection with acoustic and phonetic features for automatic speech recognition
- Subspace leakage analysis of sample data covariance matrix
- Subspace learning using consensus on the grassmannian manifold
- Subspace projection matrix completion on Grassmann manifold
- Subspace-based phase noise estimation in OFDM receivers
- Sum rate maximization model of non-regenerative multi-stream multi-pair multi-relay network
- Super-resolution acoustic imaging using sparse recovery with spatial priming
- Super-resolution in Phase Space
- Super-resolution ultrawideband ultrasound imaging using focused frequency time reversal music
- Super-wideband bandwidth extension for speech in the 3GPP EVS codec
- Supervised domain adaptation for emotion recognition from speech
- Supervised hierarchical segmentation for bird song recording
- Supervised sparse coding with local geometrical constraints
- Support knowledge-aided sparse Bayesian learning for compressed sensing
- Switching dual kernels for separable edge-preserving filtering
- Switching to and combining offline-adapted cluster acoustic models based on unsupervised segment classification
- Synchronization rules for HMM-based audio-visual laughter synthesis
- System architectures and digital signal processing algorithms for enhancing the output audio quality of stereo FM broadcast receivers
- Telephony text-prompted speaker verification using i-vector representation
- Temporal Tile Shaping for spectral gap filling in audio transform coding in EVS
- Temporal entropy-based texturedness indicator for audio signals
- Tensor object classification via multilinear discriminant analysis network
- The AMG1608 dataset for music emotion recognition
- The THUEE system for the openKWS14 keyword search evaluation
- The effect of neural networks in statistical parametric speech synthesis
- The efficiency of view synthesis prediction for 3D video coding: A spectral domain analysis
- The proportional mean decomposition: A bridge between the Gaussian and bernoulli ensembles
- The role of glottal source parameters for high-quality transformation of perceptual age
- The segregation of spatialised speech in interference by optimal mapping of diverse cues
- The shared dirichlet priors for bayesian language modeling
- The widely linear quaternion recursive total least squares
- Tikhonov-Galerkin stochastic system identification in SO(3)
- Time-frequency image descriptors-based features for EEG epileptic seizure activities detection and classification
- Time-reversal space-time codes in asynchronous two-way double-antenna relay networks
- Time-switching based SWPIT for network-coded two-way relay transmission with data rate fairness
- Time-varying vector Poisson processes with coincidences
- Token-level interpolation for class-based language models
- Tokenizing fundamental frequency variation for Mandarin tone error detection
- Tonal complexity features for style classification of classical music
- Topological interference management for two cell interference broadcast channels with alternating connectivity
- Total Jensen divergences: Definition, properties and clustering
- Total generalized variation for graph signals
- Towards machines that know when they do not know: Summary of work done at 2014 Frederick Jelinek Memorial Workshop
- Tracking changes in functional connectivity of brain networks from resting-state fMRI using particle filters
- Transient interference suppression via structured low-rank matrix decomposition
- Transmission distortion modeling for view synthesis prediction based 3-D video streaming
- Transmit code design for extended target detection in the presence of clutter
- Transmitting informative components of fisher codes for mobile visual search
- Trinicon-BSS system incorporating robust dual beamformers for noise reduction
- Twice-universal piecewise linear regression via infinite depth context trees
- Two-stage speech/music classifier with decision smoothing and sharpening in the EVS codec
- Tyler's estimator performance analysis
- Under-sampled functional MRI using low-rank plus sparse matrix decomposition
- Unicode-based graphemic systems for limited resource languages
- Unidirectional long short-term memory recurrent neural network with recurrent output layer for low-latency speech synthesis
- Unit circle MVDR beamformer
- Universal lower bounds on sampling rates for covariance estimation
- Universal outlier hypothesis testing: Application to anomaly detection
- Unnormalized exponential and neural network language models
- Unscented Transformation based array interpolation
- Unsupervised adaptation of a denoising autoencoder by Bayesian Feature Enhancement for reverberant asr under mismatch conditions
- Unsupervised data selection and word-morph mixed language model for tamil low-resource keyword search
- Unsupervised detection of malware in persistent web traffic
- Unsupervised detrending technique using sparse dictionary learning for fMRI preprocessing and analysis
- Unsupervised feature learning for urban sound classification
- Unsupervised learning of acoustic features via deep canonical correlation analysis
- Unsupervised neural network based feature extraction using weak top-down constraints
- Unsupervised speaker adaptation of deep neural network based on the combination of speaker codes and singular value decomposition for speech recognition
- Unusual event detection in crowded scenes by trajectory analysis
- Unveiling the tree: A convex framework for sparse problems
- Utilizing spectro-temporal correlations for an improved speech presence probability based noise power estimation
- Variational Bayes learning of multiscale graphical models
- Variational Bayes state space model for acoustic echo reduction and dereverberation
- Variational EM for clustering interaural phase cues in MESSL for blind source separation of speech
- Variational inference cooperative network localization with narrowband radios
- View synthesis optimization based on texture smoothness for 3D-HEVC
- Visual and acoustic identification of bird species
- Visual tracking using learned color features
- Visualization of sound field by means of Schlieren method with spatio-temporal filtering
- Vital signs from inside a helmet: A multichannel face-lead study
- Vocaine the vocoder and applications in speech synthesis
- Vocal activity informed singing voice separation with the iKala dataset
- Vocal responses to frequency modulated composite sinewaves via auditory and vibrotactile pathways
- Voice activity detection using subband noncircularity
- Voice conversion using deep Bidirectional Long Short-Term Memory based Recurrent Neural Networks
- Voice quality: Not only about "you" but also about "your interlocutor"
- Voltage sags estimation in three-phase systems using Unconditional Maximum Likelihood estimation
- WFST-based structural classification integrating dnn acoustic features and RNN language features for speech recognition
- Wave atom based Compressive Sensing and adaptive beamforming in ultrasound imaging
- Wavelet-based compressed spectrum sensing for cognitive radio wireless networks
- Weak interference direction of arrival estimation in the GPS L1 frequency band
- Weight estimation in hypergraph learning
- Weighted covariance matching based square root LASSO
- Weighted one-norm minimization with inaccurate support estimates: Sharp analysis via the null-space property
- Weighted pairwise Gaussian likelihood regression for depression score prediction
- Weighted training for speech under Lombard Effect for speaker recognition
- Wideband waveform design for robust target detection
- Wind noise short term power spectrum estimation using pitch adaptive inverse binary masks
- Wireless information and power transfer in MIMO channels under Rician fading
- Word embedding for recurrent neural network based TTS synthesis
- Word-semantic lattices for spoken language understanding
- eTutor: Online learning for personalized education
- ℓ1-constrained MVDR-based selection of nonidentical directivities in microphone array
ICASSP accepted papers in other years
Looking for submission deadlines instead? See the conference deadline calendar.