ICASSP 2016 Accepted Papers
The full list of 1,322 papers accepted at ICASSP 2016 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.
- RD-SVM: A resilient distributed support vector machine
- RSS-based sensor localization in underwater acoustic sensor networks
- Radar imaging of stationary indoor targets using joint low-rank and sparsity constraints
- Radial filters for near field source separation in spherical harmonic domain
- Radioastronomical image reconstruction with regularized least squares
- Random access for massive MIMO systems with intra-cell pilot contamination
- Random matrix based method for joint DOD and DOA estimation for large scale MIMO radar in non-Gaussian noise
- Random projections through multiple optical scattering: Approximating Kernels at the speed of light
- Randomized requantization with local differential privacy
- Range azimuth indication using a random frequency diverse array
- Rank-one tensor injection: A novel method for canonical polyadic tensor decomposition
- Ranking the parameters of deep neural networks using the fisher information
- Rate analysis for detection of sparse mixtures
- Rate analysis of spatial multiplexing in MIMO heterogeneous networks with wireless backhaul
- Rate optimization for massive MIMO relay networks: A minorization-maximization approach
- Readability enhancement of low light images based on dual-tree complex wavelet transform
- Real-time data selection and ordering for cognitive bias mitigation
- Real-time integration of statistical model-based speech enhancement with unsupervised noise PSD estimation using microphone array
- Real-time joint energy storage management and load scheduling with renewable integration
- Real-time multi-candidates fusion based head tracking on Kinect depth sequence
- Realistic human action recognition: When deep learning meets VLAD
- Recognition of occluded facial expressions based on CENTRIST features
- Reconstructing non-point sources of diffusion fields using sensor measurements
- Recovering K-sparse N-length vectors in O(K log N) time: Compressed sensing using sparse-graph codes
- Recurrent neural network training with dark knowledge transfer
- Recurrent neural networks for polyphonic sound event detection in real life recordings
- Recurrent support vector machines for speech recognition
- Recursive versions of the Levenberg-Marquardt reassigned spectrogram and of the synchrosqueezed STFT
- Reduced complexity FFT-based DOA and DOD estimation for moving target in bistatic MIMO radar
- Reduced-order modeling of hidden dynamics
- Reference-based compressed sensing: A sample complexity approach
- Region matching and similarity enhancing for image retrieval
- Regression, the periodogram, and the Lomb-Scargle periodogram
- Relative location for light field saliency detection
- Relative-gradient Bussgang-type blind equalization algorithms
- Reliably detecting humans with RGB-D camera with physical blob detector followed by learning-based filtering
- Removal of EEG artifacts for BCI applications using fully Bayesian tensor completion
- Representations of piecewise smooth signals on graphs
- Resilient decentralized consensus-based state estimation for smart grid in presence of false data
- Resource allocation for asynchronous cognitive radio networks with FBMC/OFDM under statistical CSI
- Retiming and dual-supply voltage based energy optimization for DSP applications
- Retinal vessel enhancement using multi-dictionary and sparse coding
- Retrieving audio recordings using musical themes
- Reversible data hiding in encrypted image based on block histogram shifting
- Revertible deep convolutional networks with iterated directional filter bank
- Risk assessment for RGBD scans in real time
- Risk-sensitive decision making via constrained expected returns
- Robust CDMA receiver design under disguised jamming
- Robust MVDR beamforming using time-frequency masks for online/offline ASR in noise
- Robust TTS duration modelling using DNNS
- Robust adaptive beamforming based on DOA support using decomposed coprime subarrays
- Robust artificial-noise aided transmit design for multi-user MISO systems with integrated services
- Robust audiovisual speech recognition using noise-adaptive linear discriminant analysis
- Robust blind source separation in a reverberant room based on beamforming with a large-aperture microphone array
- Robust blind spikes deconvolution
- Robust dictionary learning: Application to signal disaggregation
- Robust geographical load balancing for sustainable data centers
- Robust image hashing based on low-rank and sparse decomposition
- Robust lane marking detection using boundary-based inverse perspective mapping
- Robust multiple speech source localization using time delay histogram
- Robust pilot decontamination: A joint angle and power domain approach
- Robust pitch tracking in noisy speech using speaker-dependent deep neural networks
- Robust receiver design based on FEC code diversity in pilot-contaminated multi-user massive MIMO systems
- Robust saliency propagation based on random walks
- Robust sparse recovery for compressive sensing in impulsive noise using ℓp-norm model fitting
- Robust sparsity-promoting acoustic multi-channel equalization for speech dereverberation
- Robust speaker DOA estimation with single AVS in bispectrum domain
- Robust speech recognition from ratio masks
- Robust speech recognition using multivariate copula models
- Robust submodular data partitioning for distributed speech recognition
- Robust transmit precoding for underlay MIMO cognitive radio with interference leakage rate limit
- Robust visual tracking via inverse nonnegative matrix factorization
- Robust volume minimization-based matrix factorization via alternating optimization
- Robust waveform design of wideband cognitive radar for extended target detection
- Room geometry estimation from acoustic echoes using graph-based echo labeling
- Rotating coded aperture for depth from defocus
- SAR image target recognition using kernel sparse representation based on reconstruction coefficient energy maximization rule
- SAT-LHUC: Speaker adaptive training for learning hidden unit contributions
- SBL-based joint target imaging and Doppler frequency estimation in monostatic MIMO radar systems
- SIMD-based datapath with efficient operation structure for motion estimation
- SINR performance of matched illumination signals with dynamic target models
- SNR-invariant PLDA with multiple speaker subspaces
- STC anti-spoofing systems for the ASVspoof 2015 challenge
- SVR based double-scale regression for dynamic emotion prediction in music
- Safe screening tests for LASSO based on firmly non-expansiveness
- SalSi: A new seismic attribute for salt dome detection
- Saliency & structure preserving multi-operator image retargeting
- Saliency analysis based on depth contrast increased
- Saliency detection based on integration of central bias, reweighting and multi-scale for superpixels
- Saliency detection using tensor sparse reconstruction residual analysis
- Saliency preprocessing for person re-identification images
- Scalable training of deep learning machines by incremental block training with intra-block parallel optimization and blockwise model-update filtering
- Scaling and occlusion robust athlete tracking in sports videos
- Scanned document enhancement based on fast text detection
- Scene text recognition with high performance CNN classifier and efficient word inference
- Secrecy degrees of freedom of a MIMO Gaussian wiretap channel with a cooperative jammer
- Secure M-PSK communication via directional modulation
- Secure performance analysis of buffer-aided cognitive relay networks under delay unconstraint case
- Segment-oriented evaluation of speaker diarisation performance
- Selection and combination of hypotheses for dialectal speech recognition
- Self-stabilized deep neural network
- Semantic word embedding neural network language models for automatic speech recognition
- Semi-autonomous data enrichment based on cross-task labelling of missing targets for holistic speech analysis
- Semi-non-negative matrix factorization using alternating direction method of multipliers for voice conversion
- Semi-supervised learning in the presence of novel class instances
- Sequence design to minimize the peak sidelobe level
- Sequence summarizing neural network for speaker adaptation
- Sequence training of multi-task acoustic models using meta-state labels
- Sequential Monte Carlo sampling for correlated latent long-memory time-series
- Shadow detection using double-threshold pulse coupled neural networks
- Shape initialization without ground truth for face alignment
- Shape: Linear-time camera pose estimation with quadratic error-decay
- Shifted and convolutive source-filter non-negative matrix factorization for monaural audio source separation
- Ship wake detection for SAR images with complex backgrounds based on morphological dictionary learning
- Siamese neural network based gait recognition for human identification
- Signal detection in para complex normal noise
- Signal detection of ambient backscatter system with differential modulation
- Signal processing concepts help teach optical engineering
- Signal processing on graphs: Performance of graph structure estimation
- Signal reconstruction in the presence of side information: The impact of projection kernel design
- Signal sparsity estimation from compressive noisy projections via γ-sparsified random matrices
- Signal-adaptive switching of overlap ratio in audio transform coding
- Signer-independent fingerspelling recognition with deep neural network adaptation
- Significance of Pseudo-syllables in building better acoustic models for Indian English TTS
- Simple multi frame analysis methods for estimation of amplitude spectral envelope estimation in singing voice
- Simplified learning with binary orthogonal constraints
- Simplified multi-bit SC list decoding for polar codes
- Simplifying long short-term memory acoustic models for fast training and decoding
- Single image brightening via exposure fusion
- Single underwater image restoration by blue-green channels dehazing and red channel correction
- Single-microphone speech enhancement using MVDR filtering and Wiener post-filtering
- Sketching for large-scale learning of mixture models
- Smartphone-based real-time classification of noise signals using subband features and random forest classifier
- Smooth talking: Articulatory join costs for unit selection
- Social force model aided robust particle PHD filter for multiple human tracking
- Soft linear discriminant analysis (SLDA) for pattern recognition with ambiguous reference labels: Application to social signal processing
- Song recommendation with non-negative matrix factorization and graph total variation
- Sound field decomposition in reverberant environment using sparse and low-rank signal models
- Sound source localization based on deep neural networks with directional activate function exploiting phase information
- Source cell phone matching from speech recordings by sparse representation and KISS metric
- Source localization on solids utilizing logistic modeling of energy transition in vibration signals
- Source modeling for HMM based speech synthesis using integrated LP residual
- Source-specific system identification
- Space-shift sampling of graph signals
- Sparse Bayesian dictionary learning with a Gaussian hierarchical model
- Sparse PCA via hard thresholding for blind source separation
- Sparse attacking strategies in multi-sensor dynamic systems maximizing state estimation errors
- Sparse canonical correlation analysis based on rank-1 matrix approximation and its application for FMRI signals
- Sparse coding with fast image alignment via large displacement optical flow
- Sparse complex FxLMS for active noise cancellation over spatial regions
- Sparse deconvolution for moving-source localization
- Sparse phase retrieval with near minimal measurements: A structured sampling based approach
- Sparse reconstruction of quantized speech signals
- Sparse reconstruction-based angle-range-polarization-dependent beamforming with polarization sensitive frequency diverse array
- Sparse recovery of multiple measurement vectors in impulsive noise: A smooth block successive minimization algorithm
- Sparse signal recovery methods for variant detection in next-generation sequencing data
- Sparse sound field decomposition with multichannel extension of complex NMF
- Sparsity based multi-target tracking using mobile sensors
- Sparsity-based direction-of-arrival estimation for strictly non-circular sources
- Sparsity-based localization of spatially coherent distributed sources
- Sparsity-based reconstruction method for signals with finite rate of innovation
- Sparsity-promoting sensor selection with energy harvesting constraints
- Spatial correlation model based observation vector clustering and MVDR beamforming for meeting recognition
- Spatial feature learning for robust binaural sound source localization using a composite feature vector
- Spatio-temporal mid-level feature bank for action recognition in low quality video
- Speaker adaptation OF RNN-BLSTM for speech recognition based on speaker code
- Speaker adaptive model based on Boltzmann machine for non-parallel training in voice conversion
- Speaker adaptive training in deep neural networks using speaker dependent bottleneck features
- Speaker age estimation on conversational telephone speech using senone posterior based i-vectors
- Speaker and language factorization in DNN-based TTS synthesis
- Speaker cluster-based speaker adaptive training for deep neural network acoustic modeling
- Speaker diarization with unsupervised training framework
- Speaker recognition using matched filters
- Speaker-aware training of LSTM-RNNS for acoustic modelling
- Speech analysis of sung-speech and lyric recognition in monophonic singing
- Speech dereverberation using linear prediction with estimation of early speech spectral variance
- Speech emotion recognition using transfer non-negative matrix factorization
- Speech enhancement based on neural networks applied to cochlear implant coding strategies
- Speech enhancement using an MMSE spectral amplitude estimator based on a modulation domain Kalman filter with a Gamma prior
- Speech recognition robust against speech overlapping in monaural recordings of telephone conversations
- Spherical microphone array acoustic rake receivers
- Spoofing detection from a feature representation perspective
- Spread spectrum compressed sensing MRI using chirp radio frequency pulses
- Stability analysis of the least-mean-magnitude-phase algorithm
- Stabilization of adaptive eigenvector extraction by continuation in nested orthogonal complement structure
- Stable and symmetric filter convolutional neural network
- Stable dysphonia measures selection for Parkinson speech rehabilitation via diversity regularized ensemble
- Stacked correlation filters for biometric verification
- Statistical F0 prediction for electrolaryngeal speech enhancement considering generative process of F0 contours within product of experts framework
- Statistical analysis of neuronal population codes for encoding acute pain
- Statistical near-far detection techniques for GNSS snapshot receivers
- Steganalysis of AAC using calibrated Markov model of adjacent codebook
- Stochastic energy management in distribution grids
- Stochastic load scheduling for risk-limiting economic dispatch in smart microgrids
- Stochastic online control for smart-grid powered MIMO downlink transmissions
- Stochastic proximal gradient consensus over time-varying networks
- Stochastic thermodynamic integration: Efficient Bayesian model selection via stochastic gradient MCMC
- Structural maximum a posteriori speaker adaptation of speaking rate-dependent hierarchical prosodic model for Mandarin TTS
- Structural segmentation with the Variable Markov Oracle and boundary adjustment
- Structural spatio-temporal transform for robust visual tracking
- Structurally-constrained gradient descent for matrix factorization in haplotype assembly problems
- Structure-guided image completion via regularity statistics
- Student's T nonnegative matrix factorization and positive semidefinite tensor factorization for single-channel audio source separation
- Study of attenuation due to wet antenna in microwave radio communication
- Style retrieval from natural images
- Style-centric image summarization from photographic views of a city
- Subspace clustering with a learned dimensionality reduction projection
- Subspace fitting via sparse representation of signal covariance for DOA estimation
- Subspace superdirective beamformers based on joint diagonalization
- Subspace-based adaptive widely linear blind channel estimation for constrained minimum variance CDMA receiver
- Sum secrecy rate maximization for full-duplex two-way relay networks
- Super nested arrays: Sparse arrays with less mutual coupling than nested arrays
- Super-resolution DOA estimation via continuous group sparsity in the covariance domain
- Super-resolution spectral analysis for ultrasound scatter characterization
- Super-resolved time-of-flight sensing via FRI sampling theory
- Superimposed pilots: An alternative pilot structure to mitigate pilot contamination in massive MIMO
- Supervised and unsupervised active learning for automatic speech recognition of low-resource languages
- Supervised speech dereverberation in noisy environments using exemplar-based sparse representations
- Supervised subspace learning based on deep randomized networks
- Supervised-learning based face hallucination for enhancing face recognition
- Symmetric matrix perturbation for differentially-private principal component analysis
- Synaptic depression in deep neural networks for speech processing
- Synthesis of Volterra filters for the parametric array loudspeaker
- System architectures for communication-aware multi-robot navigation
- System combination with log-linear models
- System fusion and speaker linking for longitudinal diarization of TV shows
- System-compatible robustness improvement for new generation dect decoders by G.722 soft-decision decoding
- TC: Throughput centric successive cancellation decoder hardware implementation for polar codes
- Tag recommendation via robust probabilistic discriminative matrix factorization
- Taking meredith out of Grey's anatomy: Automating hospital ICU emergency signaling
- Target detection for depth imaging using sparse single-photon data
- Task-driven deep transfer learning for image classification
- Template based techniques for automatic segmentation of TTS unit database
- Tensor beamforming for multilinear translation invariant arrays
- Tensor completion via adaptive sampling of tensor fibers: Application to efficient indoor RF fingerprinting
- Tensor completion via functional smooth component deflation
- Tensor-based subspace learning for tracking salt-dome boundaries constrained by seismic attributes
- Terrain-scattered jammer suppression in MIMO radar using space-(fast) time adaptive processing
- Testing for impropriety of multivariate complex random processes
- Testing the consistency assumption: Pronunciation variant forced alignment in read and spontaneous speech synthesis
- The Rao test for testing handedness of complex-valued covariance matrix
- The divergence behavior of adaptive signal processing algorithms with finite search horizon
- The effect of vocal fry on pitch perception
- The first-order high-pass filter influences the automatic measurements of the electrocardiogram
- The graph FRI framework-spline wavelet theory and sampling on circulant graphs
- The intrinsic value of HFO features as a biomarker of epileptic activity
- The matching-minimization algorithm, the INCA algorithm and a mathematical framework for voice conversion with unaligned corpora
- The method for defocusing selfie taken by mobile frontal camera using burst shot
- The multiple-point variogram of images for robust texture classification
- The recursive hessian sketch for adaptive filtering
- The relationship of voice onset time and Voice Offset Time to physical age
- The spherical harmonics root-music
- The steady-state of the (Normalized) LMS is schur convex
- The use of unit norm tight measurement matrices for one-bit compressed sensing
- Theoretical guarantees for poisson disk sampling using pair correlation function
- Time domain acoustic contrast control implementation of sound zones for low-frequency input signals
- Time-resolved image demixing
- Time-varying frequency modes of resting fMRI brain networks reveal significant gender differences
- Title assignment for automatic topic segments in TV broadcast news
- Tomographic reconstruction of atmospheric density with Mumford-Shah functionals
- Towards PLDA-RBM based speaker recognition in mobile environment: Designing stacked/deep PLDA-RBM systems
- Towards a behaviorally-validated computational audiovisual saliency model
- Towards a characterization of the uncertainty curve for graphs
- Towards an automatic monitoring of the neurological state of Parkinson's patients from speech
- Towards implicit complexity control using variable-depth deep neural networks for automatic speech recognition
- Towards information-based feedback control for binaural active localization
- Towards multi-rigid body localization
- Towards optimal vlad for human action recognition from still images
- Towards robust close-talking microphone arrays for noise reduction in mobile phones
- Track selection in multifunction radars for multi-target tracking: An anti-coordination game
- Tradeoff between quality and quantity of emotional annotations to characterize expressive behaviors
- Trading accuracy for numerical stability: Orthogonalization, biorthogonalization and regularization
- Traffic-aware association in heterogeneous networks
- Training deep neural-networks based on unreliable labels
- Trajectory training considering global variance for speech synthesis based on neural networks
- Transform domain temporal prediction with extended blocks
- Transient model of EEG using Gini Index-based matching pursuit
- Tree-structured probabilistic model of monophonic written music based on the generative theory of tonal music
- Triple-based analysis of music alignments without the need of ground-truth annotations
- True time delay beamspace wideband source localization
- Turbo compressed sensing using message passing de-quantization
- Twin-HMM-based non-intrusive speech intelligibility prediction
- Two-dimensional correlated topic models
- Two-dimensional positive spline smoothing and its application to probability density estimation
- Two-stage noise aware training using asymmetric deep denoising autoencoder
- Type-2 fuzzy GMM for text-independent speaker verification under unseen noise conditions
- Typically developed adults and adults with autism spectrum disorder classification using centre of pressure measurements
- UTD-CRSS system for the NIST 2015 language recognition i-vector machine learning challenge
- Uniform expected likelihood solution for interference rejection combining regularization
- Universal encoding of multispectral images
- Universal outlying sequence detection for continuous observations
- Unmixing multitemporal hyperspectral images with variability: An online algorithm
- Unsupervised diffusion-based LMS for node-specific parameter estimation over wireless sensor networks
- Unsupervised neighbor dependent nonlinear unmixing
- Unsupervised spatiotemporal video clustering a versatile mean-shift formulation robust to total object occlusions
- Unsupervised speaker adaptation for DNN-based TTS synthesis
- Unsupervised time-series clustering of distorted and asynchronous temporal patterns
- Unsupervised user intent modeling by feature-enriched matrix factorization
- Using conditional restricted Boltzmann machines for spectral envelope modeling in speech bandwidth extension
- Using continuous lexical embeddings to improve symbolic-prosody prediction in a text-to-speech front-end
- Using hydrodynamical simulations of stellar atmospheres for periodogram standardization: Application to exoplanet detection
- VMF-SNE: Embedding for spherical data
- Variable span filters for speech enhancement
- Variational Bayesian image fusion based on combined sparse representations
- Variational inference for infinite mixtures of sparse Gaussian processes through KL-correction
- Vectorial total variation based on arranged structure tensor for multichannel image restoration
- Very deep multilingual convolutional neural networks for LVCSR
- Vibration parameter estimation using FMCW radar
- View synthesis based on temporal prediction via warped motion vector fields
- Visual tracking via multi-task non-negative matrix factorization
- Visual tracking via robust multi-task multi-feature joint sparse representation
- Visualizations relevant to the user by multi-view latent variable factorization
- Voice Morphing that improves TTS quality using an optimal dynamic frequency warping-and-weighting transform
- Wavelet features for classification of vote snore sounds
- Wavelet-based decomposition of F0 as a secondary task for DNN-based speech synthesis with multi-task learning
- What to do about noisy consensus?
- WiFi action recognition via vision-based methods
- Wide matching - An approach to improving noise robustness for speech enhancement
- Wideband multilinear array processing through tensor decomposition
- Wind speed estimation of low-altitude wind-shear based on multiple Doppler channels joint adaptive processing
- Work-efficient parallel non-maximum suppression for embedded GPU architectures
- Zero-shot learning of intent embeddings for expansion by convolutional deep structured semantic models
ICASSP accepted papers in other years
Looking for submission deadlines instead? See the conference deadline calendar.