← All conferences

ICASSP 2017 Accepted Papers

The full list of 1,320 papers accepted at ICASSP 2017 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

  1. 1+N fusion: Cascaded self-portrait enhancement
  2. 3D audio-visual speaker tracking with an adaptive particle filter
  3. 3D colored mesh graph signals multi-layer morphological enhancement
  4. 3D reconstruction from web harvested images using a forensic quality metric
  5. 3D tracking swimming fish school with learned kinematic model using LSTM network
  6. A "polyphase" structure of two-channel spectral graph wavelets and filter banks
  7. A Bayesian approach to Top-Scoring Pairs classification
  8. A Bayesian lower bound for parameter estimation of Poisson data including multiple changes
  9. A Deep Learning approach to modeling competitiveness in spoken conversations
  10. A Diagonal-Augmented quasi-Newton method with application to factorization machines
  11. A Maximum Likelihood "identification-correction" scheme of sub-optimal "SeDJoCo" solutions for semi-Blind Source Separation
  12. A Multi-resolution approach to Common Fate-based audio separation
  13. A Neural Network approach for mixing language models
  14. A PLLR and multi-stage Staircase Regression framework for speech-based emotion prediction
  15. A bayesian multi-frame image super-resolution algorithm using the Gaussian Information Filter
  16. A blind transform based approach for the detection of isolated astrophysical pulses
  17. A cache-based bandwidth optimized motion compensation architecture for video decoder
  18. A case study of machine learning hardware: Real-time source separation using Markov Random Fields via sampling-based inference
  19. A compact formulation for the l21 mixed-norm minimization problem
  20. A compact pairwise trajectory representation for action recognition
  21. A comparative study of acoustic-to-articulatory inversion for neutral and whispered speech
  22. A comparison between real and complex Schott spherical symmetry test for PolSAR data analysis
  23. A comparison of Deep Learning methods for environmental sound detection
  24. A comprehensive performance comparison of RFI mitigation techniques for UWB radar signals
  25. A comprehensive study of deep bidirectional LSTM RNNS for acoustic modeling in speech recognition
  26. A constrained adaptive scan order approach to transform coefficient entropy coding
  27. A convolutional Riemannian texture model with differential entropic active contours for unsupervised pest detection
  28. A cross-modal adaptation approach for brain decoding
  29. A data centric approach to utility change detection in online social media
  30. A data-driven compressive sensing framework tailored for energy-efficient wearable sensing
  31. A deep learning approach to multiple kernel fusion
  32. A deep learning approach towards pore extraction for high-resolution fingerprint recognition
  33. A deep neural network integrated with filterbank learning for speech recognition
  34. A deep-learning approach to translate between brain structure and functional connectivity
  35. A differentially private ensemble Kalman Filter for road traffic estimation
  36. A distributed constrained-form support vector machine
  37. A distributed robust transmit beamforming design for full-duplex relay-aided wireless communication systems
  38. A double incremental aggregated gradient method with linear convergence rate for large-scale optimization
  39. A dual estimation approach for removing the show-through effect in the scanned documents
  40. A dynamic Bayesian nonparametric model for blind calibration of sensor networks
  41. A dynamic programming approach for automatic stride detection and segmentation in acoustic emission from the knee
  42. A fast covariance matrix reconstruction method for two-dimensional direction-of-arrival estimation
  43. A fast face clustering method for indexing applications on mobile phones
  44. A fast intra-prediction decision algorithm in inter-frame based on a novel feature of HEVC
  45. A feature-based linear regression model for predicting perceptual ratings of music by cochlear implant listeners
  46. A finite rate of innovation multichannel sampling hardware system for multi-pulse signals
  47. A first attempt at polyphonic sound event detection using connectionist temporal classification
  48. A first look into a Convolutional Neural Network for speech emotion detection
  49. A full reference stereoscopic video quality assessment metric
  50. A generalization of the sparse iterative covariance-based estimator
  51. A generalized Swendsen-Wang algorithm for Bayesian nonparametric joint segmentation of multiple images
  52. A generalized log-spectral amplitude estimator for single-channel speech enhancement
  53. A generalized matrix-decomposition processor for joint MIMO transceiver design
  54. A geometric learning approach on the space of complex covariance matrices
  55. A greedy algorithm with learned statistics for sparse signal reconstruction
  56. A hierarchical Dirichlet process mixture of GID Distributions with feature selection for spatio-temporal video modeling and segmentation
  57. A joint detection-classification model for audio tagging of weakly labelled data
  58. A joint learning based Face Super Resolution approach via contextual topological structure
  59. A k-nearest neighbor multilabel ranking algorithm with application to content-based image retrieval
  60. A knowledge transfer and boosting approach to the prediction of affect in movies
  61. A knowledge-driven framework for ECG representation and interpretation for wearable applications
  62. A lattice method for resolving range ambiguity in dual-frequency RFID tag localisation
  63. A locality-preserving essence vector modeling framework for spoken document retrieval
  64. A locally linear embbeding based postfiltering approach for speech enhancement
  65. A low-complexity algorithm for utility based spectrum coordination in DSL systems
  66. A low-complexity beamforming method by orthogonal codebooks for millimeterwave links
  67. A majorization-minimization algorithm with projected gradient updates for time-domain spectrogram factorization
  68. A minimum variance partially distortionless response filter for single-channel noise reduction
  69. A mixture model-based real-time audio sources classification method
  70. A model-free causality measure based on multi-variate delay embedding
  71. A modulation feature set for robust Automatic Speech Recognition in additive noise and reverberation
  72. A multiple bandwidth objective speech intelligibility estimator based on articulation index band correlations and attention
  73. A network of deep neural networks for Distant Speech Recognition
  74. A neural filter-based scheme for synchronizing chaotic systems
  75. A neural network alternative to non-negative audio models
  76. A new chaotic feature for EEG classification based seizure diagnosis
  77. A new framework for designing incoherent sparsifying dictionaries
  78. A new generalization of the discrete Teager-Kaiser energy operator - application to biomedical signals
  79. A new noise annoyance measurement metric for urban noise sensing and evaluation
  80. A new perceptual assessment methodology for selective HEVC video encryption
  81. A new two-dimensional Fourier transform algorithm based on image sparsity
  82. A noise suppression method for body-conducted soft speech based on non-negative tensor factorization of air- and body-conducted signals
  83. A non-intrusive Short-Time Objective Intelligibility measure
  84. A nonconvex splitting method for symmetric nonnegative matrix factorization: Convergence analysis and optimality
  85. A novel LBP-based Color descriptor for face recognition
  86. A novel dictionary based SRC for face recognition
  87. A novel ensemble classifier of hyperspectral and LiDAR data using morphological features
  88. A novel iterative online rating attack based on market self-exciting property
  89. A novel layerwise pruning method for model reduction of fully connected deep neural networks
  90. A novel methodology to quantify dense EEG in cognitive tasks
  91. A novel pitch extraction based on jointly trained deep BLSTM Recurrent Neural Networks with bottleneck features
  92. A novel re-tracking strategy for monocular SLAM
  93. A novel sparse model for multi-source localization using distributed microphone array
  94. A parallelized dynamic programming approach to zero resource spoken term discovery
  95. A particle filter for sequential infection source estimation
  96. A performance-based approach to designing the stimulus presentation paradigm for the P300-based BCI by exploiting coding theory
  97. A practical high-dimensional Sparse Fourier Transform
  98. A provable nonconvex model for factoring nonnegative matrices
  99. A pseudo-Voigt component model for high-resolution recovery of constituent spectra in Raman spectroscopy
  100. A railroad detection algorithm for infrastructure surveillance using enduring airborne systems
  101. A real-time 3D head mesh modeling and expressive articulatory animation system
  102. A reassigned based singing voice pitch contour extraction method
  103. A robust FISTA-like algorithm
  104. A robust feature descriptor based on multiple gradient-related features
  105. A scalable convolutional neural network for task-specified scenarios via knowledge distillation
  106. A self-calibrating bidirectional indoor localization system
  107. A semi-supervised method for multi-subject FMRI functional alignment
  108. A simple way to approximate average robust multiuser MISO transmit optimization under covariance-based CSIT
  109. A software-defined radio implementation of timestamp-free network synchronization
  110. A sparse CCA algorithm with application to model-order selection for small sample support
  111. A speech enhancement algorithm by iterating single- and multi-microphone processing and its application to robust ASR
  112. A statistical approach to semi-supervised speech enhancement with low-order non-negative matrix factorization
  113. A stochastic maximum-likelihood framework for simplex structured matrix factorization
  114. A study of speaker verification performance with expressive speech
  115. A study on data augmentation of reverberant speech for robust speech recognition
  116. A study on motion mode identification for cyborg roaches
  117. A subspace approach for shrinkage parameter selection in undersampled configuration for Regularised Tyler Estimators
  118. A systematic approach to compute perceptual distribution of monosyllables
  119. A tensor based framework for community detection in dynamic networks
  120. A time-reversal spatial hardening effect for indoor speed estimation
  121. A transfer learning and progressive stacking approach to reducing deep model sizes with an application to speech enhancement
  122. A two-stage algorithm for noisy and reverberant speech enhancement
  123. A two-stage optimization approach to the asynchronous multi-sensor registration problem
  124. A unified convergence analysis of the multiplicative update algorithm for nonnegative matrix factorization
  125. A unified diversity measure for distributed inference
  126. A wavelet-based approach to monitoring Parkinson's disease symptoms
  127. A weakly-convex formulation for phaseless imaging
  128. ADC bit allocation under a power constraint for mmWave massive MIMO communication receivers
  129. ADMM for harmonic retrieval from one-bit sampling with time-varying thresholds
  130. AMOS: An automated model order selection algorithm for spectral graph clustering
  131. ARIMA-GARCH modeling for epileptic seizure prediction
  132. About zero bitwatermarking error exponents
  133. Accelerated dual gradient-based methods for total variation image denoising/deblurring problems
  134. Accelerated sensor position selection using graph localization operator
  135. Accelerating Deep Convolutional Networks using low-precision and sparsity
  136. Accelerating the hybrid steepest descent method for affinely constrained convex composite minimization tasks
  137. Acceleration of Adaptive normalized quasi-Newton algorithm with improved upper bounds of the condition number
  138. Achievable uplink rates for massive MIMO with coarse quantization
  139. Acoustic classification using semi-supervised Deep Neural Networks and stochastic entropy-regularization over nearest-neighbor graphs
  140. Acoustic imaging of sparse Sources with Orthogonal Matching Pursuit and clustering of basis vectors
  141. Action-vectors: Unsupervised movement modeling for action recognition
  142. Active learning for low-resource speech recognition: Impact of selection size and language modeling data
  143. Active learning for sound event classification by clustering unlabeled data
  144. Active speech control using wave-domain processing with a linear wall of dipole secondary sources
  145. Adaptation of PLDA for multi-source text-independent speaker verification
  146. Adapting and controlling DNN-based speech synthesis using input codes
  147. Adaptive DCTNet for audio signal classification
  148. Adaptive gain control and time warp for enhanced speech intelligibility under reverberation
  149. Adaptive matching pursuit for sparse signal recovery
  150. Adaptive superpixel segmentation aggregating local contour and texture features
  151. Advances in Empirical Mode Decomposition for computing Instantaneous Amplitudes and Instantaneous Frequencies
  152. Advances in all-neural speech recognition
  153. Affect recognition from lip articulations
  154. Alpha-stable multichannel audio source separation
  155. Alternating diffusion maps for dementia severity assessment
  156. Alternative networks for monolingual bottleneck features
  157. An EM algorithm for joint source separation and diarisation of multichannel convolutive speech mixtures
  158. An FFT-based synchronization approach to recognize human behaviors using STN-LFP signal
  159. An FPGA prototype of dual link algorithm for MIMO interference network
  160. An LSTM-CTC based verification system for proxy-word based OOV keyword search
  161. An M-channel critically sampled filter bank for graph signals
  162. An accumulative fusion architecture for discriminating people and vehicles using acoustic and seismic signals
  163. An accurate perturbation analysis algorithm for music with Toeplitz covariance matrix
  164. An augmented Lagrangian algorithm for decomposition of symmetric tensors of order-4
  165. An autoregressive recurrent mixture density network for parametric speech synthesis
  166. An efficient online Adaptive Sampling strategy for Matrix Completion
  167. An embedding mechanism for natural steganography after down-sampling
  168. An empirical evaluation of zero resource acoustic unit discovery
  169. An engineer's guide to Particle Filtering on the Stiefel manifold
  170. An evaluation of score-informed methods for estimating fundamental frequency and power from polyphonic audio
  171. An incremental quasi-Newton method with a local superlinear convergence rate
  172. An investigation into language model data augmentation for low-resourced STT and KWS
  173. An investigation into learning effective speaker subspaces for robust unsupervised DNN adaptation
  174. An iterative auction mechanism for data trading
  175. An iterative reconstruction algorithm for amplitude sampling
  176. An online NIPALS algorithm for Partial Least Squares
  177. An online feature selection architecture for Human Activity Recognition
  178. Analysis and prediction of heart rate using speech features from natural speech
  179. Analysis of a covert communication method utilizing non-coherent DPSK masked by pulsed radar interference
  180. Analysis of keyword spotting performance across IARPA babel languages
  181. Analytical approach to 2.5D sound field control using a circular double-layer array of fixed-directivity loudspeakers
  182. Anchor-based group detection in crowd scenes
  183. Anomaly detection in IP networks based on randomized subspace methods
  184. Anuran call classification with deep learning
  185. Aperture Domain Model Image REconstruction (ADMIRE) for improved ultrasound imaging
  186. Appearance-based gesture recognition in the compressed domain
  187. Applying compensation techniques on i-vectors extracted from short-test utterances for speaker verification using deep neural network
  188. Applying the unit circle constraint to the diagonally loaded minimum variance distortionless response beamformer
  189. Approximate simulation of linear continuous time models driven by asymmetric stable Lévy processes
  190. Array covariance matrix-based atomic norm minimization for off-grid coherent direction-of-arrival estimation
  191. Artificial bandwidth extension using the constant Q transform
  192. Assessment of broadband SNR estimation for hearing aid applications
  193. Assessment of musical noise using localization of isolated peaks in time-frequency domain
  194. Assisted dictionary learning for FMRI data analysis
  195. Asymmetric cross-view dictionary learning for person re-identification
  196. Asymptotic analysis of a GLR test for detection with large sensor arrays: New results
  197. Asymptotic analysis of multicell massive MIMO over Rician fading channels
  198. Asymptotic optimality of consensus-based sequential probability ratio test
  199. Asymptotic perfect secrecy in distributed estimation for large sensor networks
  200. Asynchronous online ADMM for consensus problems
  201. Asynchronous parallel nonconvex large-scale optimization
  202. Atlas based 3D liver segmentation using adaptive thresholding and superpixel approaches
  203. Atomic norm minimization for modal analysis with random spatial compression
  204. Audio Set: An ontology and human-labeled dataset for audio events
  205. Audio source separation based on convolutive transfer function and frequency-domain lasso optimization
  206. Audio time stretching with an adaptive multiresolution phase vocoder
  207. Audio-visual object localization and separation using low-rank and sparsity
  208. Auto-weighted two-dimensional principal component analysis with robust outliers
  209. Autoencoders trained with relevant information: Blending Shannon and Wiener's perspectives
  210. Automated robust Anuran classification by extracting elliptical feature pairs from audio spectrograms
  211. Automatic assessment of dysarthria severity level using audio descriptors
  212. Automatic conversion of Pop music into chiptunes for 8-bit pixel art
  213. Automatic detection of motion artifacts in MR images using CNNS
  214. Automatic detection of syllable stress using sonority based prominence features for pronunciation evaluation
  215. Automatic dynamic template tracking of inner lips based on CLNF
  216. Automatic gain control with integrated signal enhancement for specified target and background-noise levels
  217. Automatic image cropping with aesthetic map and gradient energy map
  218. Automatic insect recognition using optical flight dynamics modeled by kernel adaptive ARMA network
  219. Automatic matching and synchronization of user generated videos from a large scale sport event
  220. Automatic multi-lingual arousal detection from voice applied to real product testing applications
  221. Automatic musical key estimation with adaptive mode bias
  222. Automatic node selection for Deep Neural Networks using Group Lasso regularization
  223. Automatic parameter tuning for image denoising with learned sparsifying transforms
  224. Automatic radar waveform recognition based on time-frequency analysis and convolutional neural network
  225. Automatic segmentation of retinal vasculature
  226. Automatic shrinkage tuning based on a system-mismatch estimate for sparsity-aware adaptive filtering
  227. Automatic speech emotion recognition using recurrent neural networks with local attention
  228. Autoregressive moving average graph filters a stable distributed implementation
  229. Average SCR loss analysis for polarimetric STAP with Kronecker structured covariance matrix
  230. Average consensus-based asynchronous tracking
  231. Axiomatic hierarchical clustering given intervals of metric distances
  232. BER analysis of regularized least squares for BPSK recovery
  233. BLSTM-HMM hybrid system combined with sound activity detection network for polyphonic Sound Event Detection
  234. BSmCCA: A block sparse multiple-set canonical correlation analysis algorithm for multi-subject fMRI data sets
  235. Bag of Fisher Vectors representation of images by saliency-based spatial partitioning
  236. Balanced sensor management across multiple time instances via l-1/l-infinity norm minimization
  237. Balancing exploration and exploitation in reinforcement learning using a value of information criterion
  238. Barker-Coded node-pore resistive pulse sensing with built-in coincidence correction
  239. Bayesian Blind Deconvolution with application to acoustic Feedback Path modeling
  240. Bayesian information criterion for multidimensional sinusoidal order selection
  241. Bayesian joint-sequence models for grapheme-to-phoneme conversion
  242. Bayesian learning in a network with multi-hypothesis decision exchanges
  243. Bayesian multi-antenna sensing in cognitive radio networks using Fractional Bayes Factor
  244. Bayesian multichannel nonnegative matrix factorization for audio source separation and localization
  245. Bayesian nonparametric subspace estimation
  246. Bayesian phonotactic Language Model for Acoustic Unit Discovery
  247. Bayesian reconstruction of hyperspectral images by using compressed sensing measurements and a local structured prior
  248. Bayesian-driven criterion to automatically select the regularization parameter in the ℓ1-Potts model
  249. Beamnet: End-to-end training of a beamformer-supported multi-channel ASR system
  250. Belief control strategies for interactions over weak graphs
  251. Bernoulli filter based algorithm for joint target tracking and classification in a cluttered environment
  252. Binary matrix completion with performance guarantees for single individual haplotyping
  253. Biobjective transmitter optimization for service integration in MIMO Gaussian broadcast channel
  254. Biobotic motion and behavior analysis in response to directional neurostimulation
  255. Biologically inspired speech emotion recognition
  256. Bivariate probabilistic constrained programming for interference exploitation in the cognitive radio
  257. Blind bandwidth extension using K-means and Support Vector Regression
  258. Blind compensation of polynomial mixtures of Gaussian signals with application in nonlinear blind source separation
  259. Blind estimation of directional properties of room reverberation using a spherical microphone array
  260. Blind image deblurring based on sparse representation and structural self-similarity
  261. Blind image deconvolution using Student's-t prior with overlapping group sparsity
  262. Blind on board wideband antenna RF calibration for multi-antenna satellites
  263. Blind source separation based on independent low-rank matrix analysis with sparse regularization for time-series activity
  264. Blood vessels extraction using Fuzzy Mathematical Morphology
  265. Body structure based triplet Convolutional Neural Network for person re-identification
  266. Boolean Kalman Filter with correlated observation noise
  267. Borehole image correspondence and automated alignment
  268. Brain signal analytics from graph signal processing perspective
  269. Building recurrent networks by unfolding iterative thresholding for sequential sparse recovery
  270. CNN architectures for large-scale audio classification
  271. CNN-LTE: A class of 1-X pooling convolutional neural networks on label tree embeddings for audio scene classification
  272. Capacity results on the finite state Markov wiretap channel with delayed state feedback
  273. Change detection between multi-band images using a robust fusion-based approach
  274. Change detection with unknown post-change parameter using Kiefer-Wolfowitz method
  275. Channel estimation for crosstalk cancellation in wireless acoustic networks
  276. Character-level deep conflation for business data analytics
  277. Character-level language modeling with hierarchical recurrent neural networks
  278. Classification of Gaussian trajectories with missing data in Boolean gene regulatory networks
  279. Classification of thyroid nodules in ultrasound images using deep model based transfer learning and hybrid features
  280. Classification of voice modes using neck-surface accelerometer data
  281. Clinical decision support system for Parkinson's disease and related movement disorders
  282. Coalitional game theoretic optimization of electricity cost for communities of smart households
  283. Codec independent lossy audio compression detection
  284. Coding of 3D holoscopic image by using spatial correlation of rendered view images
  285. Coding of fine granular audio signals using High Resolution Envelope Processing (HREP)
  286. Coherence-adjusted monopole dictionary and convex clustering for 3D localization of mixed near-field and far-field sources
  287. Collaborative Deep Learning for speech enhancement: A run-time model selection method using autoencoders
  288. Collaborative method based on the acoustical interaction effects on active noise control systems over distributed networks
  289. Collaborative voting of 3D features for robust gesture estimation
  290. Color channel-wise recurrent learning for facial expression recognition
  291. Color demosaicking via nonlocal tensor representation
  292. Color image coding based on linear combination of adaptive colorspaces
  293. Color prediction in image coding using Steered Mixture-of-Experts
  294. Combination strategy based on relative performance monitoring for multi-stream reverberant speech recognition
  295. Combinatorial bounds on the α-divergence of univariate mixture models
  296. Combined Weighted Prediction Error and Minimum Variance Distortionless Response for dereverberation
  297. Combining belief propagation and successive cancellation list decoding of polar codes on a GPU platform
  298. Combining unidirectional long short-term memory with convolutional output layer for high-performance speech synthesis
  299. Comparison of two binaural beamforming approaches for hearing aids
  300. Complex NMF with the generalized Kullback-Leibler divergence
  301. Complexity control of HEVC for video conferencing
  302. Compressed beam-selection in millimeterwave systems with out-of-band partial support information
  303. Compressed cyclostationary detection for Cognitive Radio
  304. Compressed sensing MRI using double sparsity with additional training images
  305. Compressed sensing and optimal denoising of monotone signals
  306. Compressing higher order ambisonics of a multizone soundfield
  307. Compressive K-means
  308. Compressive imaging with iterative forward models
  309. Compressive information acquisition with hardware impairments and constraints: A case study
  310. Compressive pulse-Doppler radar sensing via 1-bit sampling with time-varying threshold
  311. Compressive sensing based ECG monitoring with effective AF detection
  312. Compressive sensing based spectrum sharing and coexistence for machine-to-machine communications
  313. Compressive sensing strategy for classification of bearing faults
  314. Computation and visualization of posterior densities in scalar nonlinear and non-Gaussian Bayesian filtering and smoothing problems
  315. Computational microscopy: illumination coding and nonlinear optimization enables Gigapixel 3D phase imaging
  316. Computing the largest eigenvalue distribution for complex Wishart matrices
  317. Concomitant of ordered multivariate normal distribution with application to parametric inference
  318. Confidence measures for CTC-based phone synchronous decoding
  319. Consensus clustering on data fragments
  320. Constrain the Docile CTUs: An In-Frame complexity allocator for HEVC Intra encoders
  321. Constructing sub-word units for spoken term detection
  322. Contextual multi-armed bandit algorithms for personalized learning action selection
  323. Contour-enhanced resampling of 3D point clouds via graphs
  324. Convergence analysis of the information matrix in Gaussian Belief Propagation
  325. Convergence rates of inertial splitting schemes for nonconvex composite optimization
  326. Convex combination framework for a priori SNR estimation in speech enhancement
  327. Convolutional Neural Network for speaker change detection in telephone speaker diarization system
  328. Convolutional approximations to linear dimensionality reduction operators
  329. Convolutional neural networks for passive monitoring of a shallow water environment using a single sensor
  330. Convolutional recurrent neural networks for music classification
  331. Copula application in nonlinear/non-Gaussian Bayesian tracking in the case of correlated sensors
  332. Correlation-based detection of TCM signals for cognitive radios
  333. Cost-effective diffusion Kalman filtering with implicit measurement exchanges
  334. Coupled hidden Markov model for automatic ECG and PCG segmentation
  335. Cover song identification with 2D Fourier Transform sequences
  336. Cramér-Rao bounds for the localization of anisotropic sources
  337. Critical sampling for wavelet filterbanks on arbitrary graphs
  338. Cross-correlations of zero crossings in jointly Gaussian and stationary processes with zero means
  339. Cross-modal transfer with neural word vectors for image feature learning
  340. Cross-modality matching based on Fisher Vector with neural word embeddings and deep image features
  341. Crowd-ML: A library for privacy-preserving machine learning on smart devices
  342. Cumulative moving averaged bottleneck speaker vectors for online speaker adaptation of CNN-based acoustic models
  343. Cyber attacks on estimation sensor networks and iots: Impact, mitigation and implications to unattacked systems
  344. D2L: Decentralized dictionary learning over dynamic networks
  345. DFVR: Deformable finger vein recognition
  346. DNN approach to speaker diarisation using speaker channels
  347. DNN-based source enhancement self-optimized by reinforcement learning using sound quality measurements
  348. DNN-based speech mask estimation for eigenvector beamforming
  349. DOA estimation in structured phase-noisy environments
  350. DOA estimation with histogram analysis of spatially constrained active intensity vectors
  351. Data analysis as a web service: A case study using IoT sensor data
  352. Data-driven fusion of multi-camera video sequences: Application to abandoned object detection
  353. Data-driven solo voice enhancement for jazz music retrieval
  354. Decentralized independent vector analysis
  355. Decoding emotional experiences through physiological signal processing
  356. Decorrelation for audio object coding
  357. Deductive refinement of species labelling in weakly labelled birdsong recordings
  358. Deep Neural Network based learning and transferring mid-level audio features for acoustic scene classification
  359. Deep attractor network for single-microphone speaker separation
  360. Deep clustering and conventional networks for music separation: Stronger together
  361. Deep fusion of heterogeneous sensor data
  362. Deep learning based automatic volume control and limiter system
  363. Deep learning on symbolic representations for large-scale heterogeneous time-series event prediction
  364. Deep long short-term memory adaptive beamforming networks for multichannel robust speech recognition
  365. Deep mixture density network for statistical model-based feature enhancement
  366. Deep multi-view models for glitch classification
  367. Deep multi-view robust representation learning
  368. Deep neural network based wake-up-word speech recognition with two-stage detection
  369. Deep neural networks based speaker modeling at different levels of phonetic granularity
  370. Deep ranking: Triplet MatchNet for music metric learning
  371. Deep salience map guided arbitrary direction scene text recognition
  372. Deep-net fusion to classify shots in concert videos
  373. DeepText: A new approach for text proposal generation and text detection in natural images
  374. Delay and Doppler processing for multi-target detection with IEEE 802.11 OFDM signaling
  375. Demixing sparse signals via convex optimization
  376. Density ridge manifold traversal
  377. Dereverberation based on bin-wise temporal variations of complex spectrogram
  378. Design of space-time block coded unique word OFDM systems
  379. Designing efficient architectures for modeling temporal features with convolutional neural networks
  380. Designing secure networks with q-composite key predistribution under different link constraints
  381. Detecting stress and depression in adults with aphasia through speech analysis
  382. Detection of Visual Evoked Potentials using Ramanujan Periodicity Transform for real time brain computer interfaces
  383. Detection of anomaly acoustic scenes based on a temporal dissimilarity model
  384. Detection of impulsive disturbances in archive audio signals
  385. Detection rate optimization in radar systems with unknown disturbance power
  386. Detection rate optimization in surveillance radars with two-step sequential detection
  387. Detection with multimodal dependent data using low-dimensional random projections
  388. Deterministic annealing based design of error resilient predictive compression systems
  389. Diagonal microphone placement for the landscape/portrait interchangeable mode of a personal computer
  390. Dialog context language modeling with recurrent neural networks
  391. Dictionary-based Equivalent Source Method for Near-Field Acoustic Holography
  392. Diffusion gradient boosting for networked learning
  393. Digital predistortion for hybrid precoding architecture in millimeter-wave massive mimo systems
  394. Direction finding using sparse linear arrays with missing data
  395. Directional discrete cosine transforms arising from discrete cosine and sine transforms for directional block-wise image representation
  396. Directional graph weight prediction for image compression
  397. Dirichlet Mixture Matching Projection for supervised linear dimensionality reduction of proportional data
  398. Dirichlet process mixture models for clustering i-vector data
  399. Disc-GLasso: Discriminative graph learning with sparsity regularization
  400. Discovering dimensions of perceived vocal expression in semi-structured, unscripted oral history accounts
  401. Discovering sound concepts and acoustic relations in text
  402. Discriminative autoencoders for speaker verification
  403. Discriminative feature domains for reverberant acoustic environments
  404. Discriminative importance weighting of augmented training data for acoustic model training
  405. Discriminative recurring signal detection and localization
  406. Disjunctive Normal Shape Boltzmann Machine
  407. Disparity estimation in stereo videos using spatio-temporal disparity hyperplane models
  408. Distance metric learning for posteriorgram based keyword search
  409. Distance-preserving property of random projection for subspaces
  410. Distributed TV-L1 image fusion using PDMM
  411. Distributed blind equalization in networked systems
  412. Distributed decision-making over mobile adaptive networks
  413. Distributed largest eigenvalue detection
  414. Distributed max-SINR speech enhancement with ad hoc microphone arrays
  415. Distributed nonconvex optimization for sparse representation
  416. Distributed optimization for evolving networks of growing connectivity
  417. Distributed probabilistic bisection search using social learning
  418. Distributed recursive least-squares with data-adaptive censoring
  419. Distributed sensor selection for field estimation
  420. Distributed sparsified graph filters for denoising and diffusion tasks
  421. Divide-and-warp temporal alignment of speech signals between speakers: Validation using articulatory data
  422. DoF analysis in a two-layered heterogeneous wireless interference network
  423. Domain adaptation of DNN acoustic models using knowledge distillation
  424. Double Relay Communication Protocol with power control for achieving fairness in cellular systems
  425. Double-bit quantization and weighting for nearest neighbor search
  426. Drum extraction in single channel audio signals using multi-layer Non negative Matrix Factor Deconvolution
  427. Drum transcription from polyphonic music with recurrent neural networks
  428. Dual-Tree wavelet scattering network with parametric log transformation for object classification
  429. Dual-fisheye lens stitching for 360-degree imaging
  430. Duration prediction using multiple Gaussian process experts for GPR-based speech synthesis
  431. Dynamic Graph Fourier Transform on temporal functional connectivity networks
  432. Dynamic Probabilistic Linear Discriminant Analysis for video classification
  433. Dynamic cloud Offloading for View Synthesis
  434. Dynamic polygon cloud compression
  435. Dynamic reconstruction of influence graphs with adaptive directed information
  436. Dynamic tracking attention model for action recognition
  437. ECG-based biometrics using recurrent neural networks
  438. EEG channel optimization via sparse common spatial filter
  439. EEG source imaging assists decoding in a face recognition task
  440. Edge-preserving filtering by projection onto L0 gradient constraint
  441. Edited film alignment via selective Hough transform and accurate template matching
  442. Effect of acoustic conditions on algorithms to detect Parkinson's disease from speech
  443. Effect of sampling on the estimation of the apparent coefficient of diffusion in MRI
  444. Effective Fisher vector aggregation for 3D object retrieval
  445. Effective articulatory modeling for pronunciation error detection of L2 learner without non-native training data
  446. Effective compressive sensing via reweighted total variation and weighted nuclear norm regularization
  447. Effective emotion recognition in movie audio tracks
  448. Effective estimation of the desired-signal subspace and its application to robust adaptive beamforming
  449. Effective joint training of denoising feature space transforms and Neural Network based acoustic models
  450. Effective keyword search for low-resourced conversational speech
  451. Effects of gender information in text-independent and text-dependent speaker verification
  452. Efficient adaptive filtering in compressive domains for sparse systems and relation to transform-domain adaptive filtering
  453. Efficient bridging-based destination inference in object tracking
  454. Efficient hybrid space-ground precoding techniques for multi-beam satellite systems
  455. Efficient large scale antenna selection by partial switching connectivity
  456. Efficient methods to train multilingual bottleneck feature extractors for low resource keyword search
  457. Efficient mode decision for noisy video transcoding
  458. Efficient multidimensional parameter estimation for joint wideband radar and communication systems based on OFDM
  459. Efficient multiplier-less structures for Ramanujan filter banks
  460. Efficient pooling of image based CNN features for action recognition in videos
  461. Efficient postcoding filter in LU-based beamforming scheme
  462. Efficient representation of segmentation contours using chain codes
  463. Efficient single/multiple unimodular waveform design with low weighted correlations
  464. Eigenvalue decomposition based estimators of carrier frequency offset in multicarrier underwater acoustic communication
  465. Embedded clustering via robust orthogonal least square discriminant analysis
  466. Emitter source localization using time-of-arrival measurements from single moving receiver
  467. Emotion estimation via tensor-based supervised decision-level fusion from multiple Brodmann areas
  468. Emotion recognition through integrating EEG and peripheral signals
  469. Encoder-decoder with focus-mechanism for sequence labelling based spoken language understanding
  470. End-to-end ASR-free keyword search from speech
  471. End-to-end joint learning of natural language understanding and dialogue manager
  472. End-to-end speech recognition and keyword search on low-resource languages
  473. End-to-end spoofing detection with raw waveform CLDNNS
  474. End-to-end visual speech recognition with LSTMS
  475. Energy blowup for truncated stable LTI systems
  476. Energy reduction opportunities in an HEVC real-time encoder
  477. Energy-efficient design for non-regenerative MIMO relay networks
  478. Engagement detection for children with Autism Spectrum Disorder
  479. Enhanced LBP texture features from time frequency representations for acoustic scene classification
  480. Enhanced canonical correlation analysis with local density for cross-domain visual classification
  481. Enhanced depth estimation for hand-held light field cameras
  482. Enhanced indoor localization through crowd sensing
  483. Enhanced pixel-wise voting for image vanishing point detection in road scenes
  484. Enhanced single antenna interference cancellation from MMSE third-order complex Volterra filters
  485. Enhanced ultrasound image reconstruction using a compressive blind deconvolution approach
  486. Enhancing ICA performance by exploiting sparsity: Application to FMRI analysis
  487. Enhancing QoS in spatially controlled beamforming networks via distributed stochastic programming
  488. Enhancing noise and pitch robustness of children's ASR
  489. Enhancing observability in power distribution grids
  490. Enhancing retinal vessel segmentation by color fusion
  491. Enhancing utility and privacy with noisy minimax filters
  492. Ensemble classification based on Random linear base classifiers
  493. Ensemble feature selection for domain adaptation in speech emotion recognition
  494. Environment aware speaker diarization for moving targets using parallel DNN-based recognizers
  495. Epithelium-stroma classification in histopathological images via convolutional neural networks and self-taught learning
  496. Estimating sparse signals using integrated wide-band dictionaries
  497. Estimation accuracy of non-standard maximum likelihood estimators
  498. Estimation and learning of Dynamic Nonlinear Networks (DyNNets)
  499. Estimation in autoregressive processes with partial observations
  500. Estimation of multiple pitches in stereophonic mixtures using a codebook-based approach
  501. Estimation of vocal tract area function from volumetric Magnetic Resonance Imaging
  502. Evaluating automatic speech recognition systems in comparison with human perception results using distinctive feature measures
  503. Evaluation of a complementary hearing aid for spatial sound segregation
  504. Evaluation of weight sparsity regularizion schemes of deep neural networks applied to functional neuroimaging data
  505. Event-based consensus for a class of heterogeneous multi-agent systems: An LMI approach
  506. Event-related synchronisation responses to N-back memory tasks discriminate between healthy ageing, mild cognitive impairment, and mild Alzheimer's disease
  507. Evolutionary affinity propagation
  508. Example-based Visual Object Counting for complex background with a local low-rank constraint
  509. Exemplar selection methods in voice conversion
  510. Exemplar-based image completion via new quality measure based on phaseless texture features
  511. Exemplar-embed complex matrix factorization for facial expression recognition
  512. Expected Likelihood sphericity test distribution for complex angular central Gaussian data
  513. Experimental demonstration of nullforming from a fully wireless distributed array
  514. Exploiting different word clusterings for class-based RNN language modeling in speech recognition
  515. Exploiting mutual coupling by means of analog-digital zero forcing
  516. Exploiting sequence information for text-dependent Speaker Verification
  517. Exploiting sequential Low-Rank Factorization for multilingual DNNS
  518. Exploring universal speech attributes for speaker verification
  519. Expressive visual text to speech and expression adaptation using deep neural networks
  520. Extended Kalman filter for extended object tracking
  521. Extended low-rank plus diagonal adaptation for deep and recurrent neural networks
  522. Extracting Fourier descriptors from compressive measurements
  523. Extracting structural spectral features using what-where auto-encoders for statistical parametric speech synthesis
  524. Extraction of common task signals and spatial maps from group fMRI using a PARAFAC-based tensor decomposition technique
  525. Extreme image completion
  526. FRI sampling and time-varying pulses: Some theory and four short stories
  527. FRIDA: FRI-based DOA estimation for arbitrary array layouts
  528. Face Album: Towards automatic photo management based on person identity on mobile phones
  529. Face detection and recognition for home service robots with end-to-end deep neural networks
  530. Face recognition in real-world images
  531. Facial attractiveness prediction using psychologically inspired convolutional neural network (PI-CNN)
  532. Factor analysis methods for joint speaker verification and spoof detection
  533. Fast HEVC intra coding algorithm based on machine learning and Laplacian Transparent Composite Model
  534. Fast HRFT measurement system with unconstrained head movements for 3D audio in virtual and augmented reality applications
  535. Fast Spectral Clustering with efficient large graph construction
  536. Fast algorithm for statistical phrase/accent command estimation based on generative model incorporating spectral features
  537. Fast and privacy preserving distributed low-rank regression
  538. Fast camera self-calibration for synthesizing Free Viewpoint soccer Video
  539. Fast convolutional sparse coding with separable filters
  540. Fast exemplar selection algorithm for matrix approximation and representation: A variant oASIS algorithm
  541. Fast feasibility pursuit for non-convex QCQPS via first-order methods
  542. Fast harmonic chirp summation
  543. Fast human segmentation using color and depth
  544. Fast hyperspectral unmixing in presence of sparse multiple scattering nonlinearities
  545. Fast implementation for symmetric non-separable transforms based on grids
  546. Fast interpolation of bandlimited functions
  547. Fast inverse tone mapping with Reinhard's global operator
  548. Fast orthogonal approximations of sampled sinusoids and bandlimited signals
  549. Fast path localization on graphs via multiscale Viterbi decoding
  550. Fast sparse recovery for any RIP-1 matrix
  551. Fast tagging of natural sounds using marginal co-regularization
  552. Faster sequence training
  553. Faster-than-Nyquist spatiotemporal symbol-level precoding in the downlink of multiuser MISO channels
  554. Feature encoding in band-limited distributed surveillance systems
  555. Feature extraction using multimodal convolutional neural networks for visual speech recognition
  556. Feature mapping for speaker diarization in noisy conditions
  557. Feature++: Cross dimension feature fusion for road detection
  558. Feature-based ROI generation for stereo-based pedestrian detection
  559. Feedback connection for deep neural network-based acoustic modeling
  560. Fetal heart rate classification by non-parametric Bayesian methods
  561. Filter design for delay-based anonymous communications
  562. First-person action recognition through Visual Rhythm texture description
  563. Fixed-point optimization of deep neural networks with adaptive step size retraining
  564. Flat focus: depth of field analysis for the FlatCam lensless imaging system
  565. Flexarray: Random phased array layouts for analytical spatial filtering
  566. Flexible large-scale fMRI analysis: A survey
  567. Flipflop correlation tracking with Convolution Kernels Networks
  568. Flow based botnet detection through semi-supervised active learning
  569. Forecasting covariance for optimal carry trade portfolio allocations
  570. Frequency-domain under-modelled blind system identification based on cross power spectrum and sparsity regularization
  571. Frequency-tuned ACM for biomedical image segmentation
  572. Frequency-warped time-weighted linear prediction for glottal vocoding
  573. From biomedical imaging to urban data mining: Theory of signal representations
  574. From focal stacks to tensor display: A method for light field visualization without multi-view images
  575. From image quality to patch quality: An Image-Patch Model for No-Reference image quality assessment
  576. Full-duplex relaying under I/Q imbalance using improper Gaussian signaling
  577. Full-duplex self-interference mitigation analysis for direct conversion RF nonlinear MIMO channel models with IQ mismatch
  578. Fully adaptive mode decomposition from time-frequency ridges
  579. Fully complex deep neural network for phase-incorporating monaural source separation
  580. Fused estimation of sparse connectivity patterns from rest fMRI
  581. Fusing shallow and deep learning for bioacoustic bird species classification
  582. Fusing structure from motion and lidar for dense accurate depth map estimation
  583. Fusing transcription results from polyphonic and monophonic audio for singing melody transcription in polyphonic music
  584. Fusion of multiple emotion perspectives: Improving affect recognition through integrating cross-lingual emotion information
  585. GDspike: An accurate spike estimation algorithm from noisy calcium fluorescence signals
  586. Game theoretic resource allocation form-dependent channels with application to OFDMA
  587. General scale interpolation via context-aware autoregressive model and multiplanar constraint
  588. Generalization of spoofing countermeasures: A case study with ASVspoof 2015 and BTAS 2016 corpora
  589. Generalized Barankin-type lower bounds for misspecified models
  590. Generalized Linear Models for count time series
  591. Generative adversarial network-based postfilter for statistical parametric speech synthesis
  592. Geometry-adapted Gaussian random field regression
  593. Global behavior of parallel projection method for certain nonconvex feasibility problems
  594. Globally optimal beamforming design for downlink CoMP transmission with limited backhaul capacity
  595. Good features to track for RGBD images
  596. Gradient magnitude similarity deviation on multiple scales for color image quality assessment
  597. Gradient-based solution for hybrid precoding in MIMO systems
  598. Graph Fourier Transform for directed graphs based on Lovász extension of min-cut
  599. Graph learning under sparsity priors
  600. Graph regularised tensor factorisation of EEG signals based on network connectivity measures
  601. Graph-signal reconstruction and blind deconvolution for diffused sparse inputs
  602. Grasp: A matlab toolbox for graph signal processing
  603. Greedy alternative for room geometry estimation from acoustic echoes: A subspace-based method
  604. Greedy search for descriptive spatial face features
  605. Gridless compressed sensing under shift-invariant sampling
  606. Group-level support recovery guarantees for group lasso estimator
  607. Guided deep network for depth map super-resolution: How much can color help?
  608. HEVC-based motion compensated joint temporal-spatial video denoising
  609. Hand pose recognition in First Person Vision through graph spectral analysis
  610. Hardware and software for reproducible research in audio array signal processing
  611. Hardware-based linear programming decoding via the alternating direction method of multipliers
  612. Harmonic feature fusion for robust neural network-based acoustic modeling
  613. Harmonic minimum mean squared error filters for multichannel speech enhancement
  614. Harnessing neural networks: A random matrix approach
  615. Hearing in a shoe-box: Binaural source position and wall absorption estimation using virtually supervised learning
  616. Heartmate: automated integrated anomaly analysis for effective remote cardiac health management
  617. Heuristic methods for designing unimodular code sequences with performance guarantees
  618. Hierarchical Structured Dictionary Learning for image categorization
  619. Hierarchical joint-guided networks for semantic image segmentation
  620. Hierarchical saliency optimization
  621. High accuracy event detection for Non-Intrusive Load Monitoring
  622. High dimensional decomposition of coherent/structured matrices via sequential column/row sampling
  623. High frequency moments via max-stability
  624. High level synthesis of Smith-Waterman dataflow implementations
  625. High precision robust modeling of long room responses using wavelet transform
  626. High-level synthesis implementation of HEVC 2-D DCT/DST on FPGA
  627. High-resolution Direction-of-Arrival estimation in SNR and snapshot challenged scenarios using multi-frequency coprime arrays
  628. Homography-based low rank approximation of light fields for compression
  629. How little does non-exact recovery help in group testing?
  630. How should we evaluate supervised hashing?
  631. Human action recognition using Adaptive Hierarchical Depth Motion Maps and Gabor filter
  632. Human interaction recognition using low-rank matrix approximation and super descriptor tensor decomposition
  633. Human recognition from photoplethysmography (PPG) based on non-fiducial features
  634. Hybrid beamforming design with finite-resolution phase-shifters for frequency selective massive MIMO channels
  635. Hybrid beamforming for large-scale MIMO systems using uplink-downlink duality
  636. Hybrid beamforming in uplink massive MIMO systems in the presence of blockers
  637. Hybrid precoding using long-term channel statistics for massive MIMO systems
  638. Hyperarticulation detection in repetitive voice queries using pairwise comparison for improved speech recognition
  639. Hyperspectral image restoration by Hybrid Spatio-Spectral Total Variation
  640. Hyperspectral unmixing with endmember variability using Partial Membership Latent Dirichlet Allocation
  641. Hypothesis testing in the presence of maxwell's daemon: signal detection by unlabeled observations
  642. ICA based single microphone Blind Speech Separation technique using non-linear estimation of speech
  643. Identifying FMRI dynamic connectivity states using affinity propagation clustering method: Application to schizophrenia
  644. Identifying a multiple plane plenoptic function from a swiped image
  645. Identifying correlated components in high-dimensional multivariate Gaussian models
  646. Identifying directional connections in brain networks via multi-kernel granger models
  647. Illumination-robust face recognition with Block-based Local Contrast Patterns
  648. Image classification: A hierarchical dictionary learning approach
  649. Image co-saliency detection via locally adaptive saliency map fusion
  650. Image compression with Stochastic Winner-Take-All Auto-Encoder
  651. Image denoising via collaborative support-agnostic recovery
  652. Image denoising via group sparsity residual constraint
  653. Image formation methods in quantitative acoustic microscopy
  654. Image recognition based on discriminative models using features generated from separable lattice HMMS
  655. Image reconstruction from partial Fourier measurements via curl constrained sparse gradient estimation
  656. Image retrieval based on deep Convolutional Neural Networks and binary hashing learning
  657. Impact of low-precision deep regression networks on single-channel source separation
  658. Implementation of efficient, low power deep neural networks on next-generation intel client platforms
  659. Implementation strategies of the seismic Full Waveform Inversion
  660. Improved Local Spectral Unmixing of hyperspectral data using an algorithmic regularization path for collaborative sparse regression
  661. Improved cepstra minimum-mean-square-error noise reduction algorithm for robust speech recognition
  662. Improved eigenvalue shrinkage using weighted Chebyshev polynomial approximation
  663. Improved template based chord recognition using the CRP feature
  664. Improving audio-visual speech recognition using deep neural networks with dynamic stream reliability estimates
  665. Improving latency-controlled BLSTM acoustic models for online speech recognition
  666. Improving mesh-based motion compensation by using edge adaptive graph-based compensated wavelet lifting for medical data sets
  667. Improving music source separation based on deep neural networks through data augmentation and network blending
  668. Improving the perceptual quality of ideal binary masked speech
  669. Improving the spatial dimensionality of Gauss-Legendre and equiangular sampling schemes on the sphere
  670. In-situ calibration of accelerometers in body-worn sensors using quiescent gravity
  671. Inband full-duplex radio access system with self-backhauling: transmit power minimization under QOS requirements
  672. Incident field recovery for an arbitrary-shaped scatterer
  673. Incremental adaptation using active learning for acoustic emotion recognition
  674. Indoor mapping using MIMO radio channel measurements
  675. Indoor multi-sound source localization based on nonparametric Bayesian clustering
  676. Induced bias in attenuation measurements taken from commercial microwave links
  677. Inference Machines for supervised Bluetooth localization
  678. Inferring emotions from heterogeneous social media data: A Cross-media Auto-Encoder solution
  679. Inferring latent states in a network influenced by neighbor activities: An undirected generative approach
  680. Inferring sparse graphs from smooth signals with theoretical guarantees
  681. Infinite-dimensional SVD for analyzing microphone array
  682. Infomax-ICA using Hessian-free optimization
  683. Information diffusion in interconnected heterogeneous networks
  684. Information geometry metric for random signal detection in large random sensing systems
  685. Information theoretic structure learning with confidence
  686. Informed source separation via compressive graph signal sampling
  687. Infrasonic scene fingerprinting for authenticating speaker location
  688. Infrastructure-less indoor localization using light fingerprints
  689. Inpainting-based error concealment for low-delay video communication
  690. Integrated DNN-based model adaptation technique for noise-robust speech recognition
  691. Integrating DNN-based and spatial clustering-based mask estimation for robust MVDR beamforming
  692. Integration of multiple genomic imaging data for the study of schizophrenia using joint nonnegative matrix factorization
  693. Intelligent compressive data gathering using data ferries for wireless sensor networks
  694. Inter dataset variability modeling for speaker recognition
  695. Inter-block dependencies consideration for intra coding in H.264/AVC and HEVC standards
  696. Interaural time delay personalisation using incomplete head scans
  697. Interference alignment on MIMO X channel with synergistic CSIT
  698. Interference cancellation in two-channel nuclear quadrupole resonance measurements
  699. Interference reduction in music recordings combining Kernel Additive Modelling and Non-Negative Matrix Factorization
  700. Interpretable human action recognition in compressed domain
  701. Interpretable phonological features for clinical applications
  702. Intra-class covariance adaptation in PLDA back-ends for speaker verification
  703. Introducing complex functional link polynomial filters
  704. Investigations on byte-level convolutional neural networks for language modeling in low resource speech recognition
  705. Iterative beam alignment algorithms for TDD MIMO systems
  706. Iterative block tensor singular value thresholding for extraction of lowrank component of image data
  707. Iterative diffusion-based anomaly detection
  708. Jamming Massive MIMO using Massive MIMO: Asymptotic separability results
  709. Jamming resistant receivers for massive MIMO
  710. Jazz: A companion to music for frequency estimation with missing data
  711. Jeffrey's divergence between moving-average and autoregressive models
  712. Joint Bayesian Gaussian Discriminant Analysis for speaker verification
  713. Joint CTC-attention based end-to-end speech recognition using multi-task learning
  714. Joint Near-End Listening Enhancement and far-end noise reduction
  715. Joint alpha-fairness based DSM and user encoding ordering for zero-forcing nonlinear precoding in G. fast downstream transmission
  716. Joint analog and digital self-interference cancellation and full-duplex system performance
  717. Joint channel and carrier frequency estimation for M-ary CPM over frequency-selective channel using PAM decomposition
  718. Joint modeling of articulatory and acoustic spaces for continuous speech recognition tasks
  719. Joint optimisation of tandem systems using Gaussian mixture density neural network discriminative sequence training
  720. Joint parameter and state estimation for wave-based imaging and inversion
  721. Joint power and subcarrier allocation for multicarrier full-duplex systems
  722. Joint transmit beamforming optimization and uplink/downlink user selection in a full-duplex multi-user MIMO system
  723. Jointly optimized transform domain temporal prediction and sub-pixel interpolation
  724. Kalman filter based system identification exploiting the decorrelation effects of linear prediction
  725. Kernel least mean square based on conjugate gradient
  726. Kernel principal component analysis of the ear morphology
  727. Kernel weighted Fisher sparse analysis on multiple maps for audio event recognition
  728. Key frames extraction using graph modularity clustering for efficient video summarization
  729. Knowledge distillation across ensembles of multilingual models for low-resource languages
  730. Knowledge distillation for small-footprint highway networks
  731. LBP edge-mapped descriptor using MGM interest points for face recognition
  732. LDA-based context dependent recurrent neural network language model using document-based topic distribution of words
  733. LDPC code design for Gaussian multiple-access channels using dynamic EXIT chart analysis
  734. LPCV: Learning projections from corresponding views for person re-identification
  735. Laplace gradient based Discriminative and Contrast Invertible descriptor
  736. Laplace mixtures models for efficient compressed sensing with side information
  737. Large scale 2D spectral compressed sensing in continuous domain
  738. Large-scale audio event discovery in one million YouTube videos
  739. Large-scale nonconvex stochastic optimization by Doubly Stochastic Successive Convex approximation
  740. Largest center-specific margin for dimension reduction
  741. Late reverberant power spectral density estimation based on an eigenvalue decomposition
  742. Latent tree approximation in linear model
  743. Learning Grassmann manifolds for object state discovery
  744. Learning a hierarchical spatio-temporal model for human activity recognition
  745. Learning and free energies for vector approximate message passing
  746. Learning and inferring human actions with temporal pyramid features based on conditional random fields
  747. Learning by networked agents under partial information
  748. Learning complex-valued latent filters with absolute cosine similarity
  749. Learning concepts through conversations in spoken dialogue systems
  750. Learning conditional independence structure for high-dimensional uncorrelated vector processes
  751. Learning cross-lingual knowledge with multilingual BLSTM for emphasis detection with limited training data
  752. Learning deep vector regression model for no-reference image quality assessment
  753. Learning dictionary for efficient signal compression
  754. Learning discriminative features from electroencephalography recordings by encoding similarity constraints
  755. Learning environmental sounds with end-to-end convolutional neural network
  756. Learning online alignments with continuous rewards policy gradient
  757. Learning representations of emotional speech with deep convolutional generative adversarial networks
  758. Learning rotation invariance in deep hierarchies using circular symmetric filters
  759. Learning sparse graphs under smoothness prior
  760. Learning spectrum opportunities in non-stationary radio environments
  761. Learning time varying graphs
  762. Learning to invert: Signal recovery via Deep Convolutional Networks
  763. Learning utterance-level representations for speech emotion and age/gender recognition using deep neural networks
  764. Least 1-norm pole-zero modeling with sparse deconvolution for speech analysis
  765. Leveraging manifold learning for extractive broadcast news summarization
  766. Line detection in speckle images using Radon transform and ℓ1 regularization
  767. Linear Discriminant Analysis with few training data
  768. Linear demixed domain multichannel nonnegative matrix factorization for speech enhancement
  769. Linear systems approach to identifying performance bounds in indirect imaging
  770. Listening-area-informed sound field reproduction based on circular harmonic expansion
  771. Local detection and estimation of multiple objects from images with overlapping observation areas
  772. Local trilateral upsampling for thermal image
  773. Locality Sensitive Hashing based deepmatching for optical flow estimation
  774. Localization of multiple sources using time-difference of arrival measurements
  775. Locally linear embedded sparse coding for image representation
  776. Location-aware network operation for cloud radio access network
  777. LogNet: Energy-efficient neural networks using logarithmic computation
  778. Lombard speech synthesis using long short-term memory recurrent neural networks
  779. Long-term non-contact tracking of caged rodents
  780. Low Dimensional Deep Features for facial landmark alignment
  781. Low angle direction of arrival estimation by time reversal
  782. Low light image enhancement based on two-step noise suppression
  783. Low rank phase retrieval
  784. Low-complexity optimization for two-dimensional direction-of-arrival estimation via decoupled atomic norm minimization
  785. Low-latency real-time blind source separation for hearing aids based on time-domain implementation of online independent vector analysis with truncation of non-causal components
  786. Low-rank and sparse soft targets to learn better DNN acoustic models
  787. Low-rank physical model recovery from low-rank signal approximation
  788. Low-resource grapheme-to-phoneme conversion using recurrent neural networks
  789. Lyric recognition in monophonic singing using pitch-dependent DNN
  790. Machine learning based non-intrusive quality estimation with an augmented feature set
  791. Making and gaming in signal processing classes
  792. Malware classification with LSTM and GRU language models and a character-level CNN
  793. Maritime anomaly detection in ferry tracks
  794. Massive MIMO processing at the semiconductor edge: Exploiting the system and circuit margins for power savings
  795. Massive device activity detection by approximate message passing
  796. Matched subspace detection using compressively sampled data
  797. Matrix completion based MIMO radars with clutter and interference mitigation via transmit precoding
  798. Matrix completion of noisy graph signals via proximal gradient minimization
  799. Maximum secrecy rate in inhomogeneous poisson networks
  800. Measurement of 2D vibration modes using amplification of high speed video in the presence of noise
  801. Measurement of sound fields using moving microphones
  802. Measuring, modelling and predicting perceived reverberation
  803. Meeting different QoS requirements of vehicular networks: A D2D-based approach
  804. Melody extraction and detection through LSTM-RNN with harmonic sum loss
  805. Memory visualization for gated recurrent neural networks in speech recognition
  806. Minimum Bayes risk training of CTC acoustic models in maximum a posteriori based decoding framework
  807. Minimum entropy pursuit: Noise analysis
  808. Minimum mean square deviation in ZA-NLMS algorithm
  809. Minimum number of possibly non-contiguous samples to distinguish two periods
  810. Minimum precision requirements for the SVM-SGD learning algorithm
  811. Minimum probability-of-error perturbation precoding for the one-bit massive MIMO downlink
  812. Mismatched sparse denoiser requires overestimating the support length
  813. Mixture source identification in non-stationary data streams with applications in compression
  814. Mobile phone clustering from acquired speech recordings using deep Gaussian supervector and spectral clustering
  815. Model based binaural enhancement of voiced and unvoiced speech
  816. Model order selection for sampling FRI signals
  817. Modeling Sallen-Key audio filters in the Wave Digital domain
  818. Modeling interest-based social networks: Superimposing Erdős-Rényi graphs over random intersection graphs
  819. Modification on LSA speech enhancement for speech recognition
  820. Modified nonnegative matrix factorization for endmember spectra extraction from highly mixed hyperspectral images combined with multispectral data
  821. Monte Carlo exploration for active binaural localization
  822. Mood detection from daily conversational speech using denoising autoencoder and LSTM
  823. Morph-to-word transduction for accurate and efficient automatic speech recognition and keyword search
  824. Motion clustering with hybrid-sample-based foreground segmentation for moving cameras
  825. Motion compensated frame rate up-conversion using 3D frequency selective extrapolation and a multi-layer consistency check
  826. Motion informed audio source separation
  827. Moving target localization in multistatic sonar using time delays, Doppler shifts and arrival angles
  828. Multi-accent speech recognition with hierarchical grapheme based models
  829. Multi-armed bandits in multi-agent networks
  830. Multi-channel noise reduction for hands-free voice communication on mobile phones
  831. Multi-channel signal enhancement with speech and noise covariance estimates computed by a probabilistic localization model
  832. Multi-pitch estimation using semidefinite programming
  833. Multi-pitch streaming of interwoven streams
  834. Multi-rate polar codes for solid state drives
  835. Multi-scale higher order singular value decomposition (MS-HoSVD) for resting-state FMRI compression and analysis
  836. Multi-scale spot segmentation with selection of image scales
  837. Multi-speaker conversations, cross-talk, and diarization for speaker recognition
  838. Multi-speaker voice activity detection by an improved multiplicative non-negative independent component analysis with sparseness constraints
  839. Multi-task deep neural network with shared hidden layers: Breaking down the wall between emotion representations
  840. Multi-task learning for face identification and attribute estimation
  841. Multi-task learning of structured output layer bidirectional LSTMS for speech synthesis
  842. Multi-view representation learning via gcca for multimodal analysis of Parkinson's disease
  843. Multichannel audio source separation: Variational inference of time-frequency sources from time-domain observations
  844. Multicore distributed dictionary learning: A microarray gene expression biclustering case study
  845. Multilayer sensor network for information privacy
  846. Multimodal sparse Bayesian dictionary learning applied to multimodal data classification
  847. Multiple illumination phaseless super-resolution (MIPS) with applications to phaseless DoA estimation and diffraction imaging
  848. Multiple parallel branch with folding architecture for multichannel filtered-x least mean square algorithm
  849. Multiple particle filtering for inference in the presence of state correlation of unknown mixing parameters
  850. Multiple sound source localization based on TDOA clustering and multi-path matching pursuit
  851. Multiple source localization using Estimation Consistency in the Time-Frequency domain
  852. Multiple subspace matching pursuit for spectrum sensing
  853. Multiple wavelength sensing array design
  854. Multiple-input multiple-output (MIMO) MRI: An efficient pulse design algorithm to combine parallel excitation and parallel imaging
  855. Multiprocessor approximate message passing with column-wise partitioning
  856. Multisensor detection of improper signals in improper noise
  857. Multistream quickest change detection: Asymptotic optimality under a sparse signal
  858. Multisymbol with memory noncoherent detection of CPFSK
  859. Multivariate Linear Time-Frequency modeling and adaptive robust target detection in highly textured monovariate SAR image
  860. Multivariate Scale mixtures for joint sparse regularization in multi-task learning
  861. Multivariate scale-free dynamics: Testing fractal connectivity
  862. Music staging AI
  863. NAPLib: An open source toolbox for real-time and offline Neural Acoustic Processing
  864. NIQSV: A no reference image quality assessment metric for 3D synthesized views
  865. Near-optimal sample complexity bounds for circulant binary embedding
  866. Nesterov-based parallel algorithm for large-scale nonnegative tensor factorization
  867. Network architectures for multilingual speech representation learning
  868. Network discovery using content and homophily
  869. Network topology inference from non-stationary graph signals
  870. Network-based genome wide study of hippocampal imaging phenotype in Alzheimer's Disease to identify functional interaction modules
  871. Neural decoding systems using Markov Decision Processes
  872. New analysis of radar micro-Doppler gait signatures for rehabilitation and assisted living
  873. New asymptotic properties for the robust ANMF
  874. New residue arithmetic based Barrett algorithms: Modular polynomial computations
  875. Node embedding for network community discovery
  876. Noise detection in smartphone phonocardiogram
  877. Noise enhanced distributed Bayesian estimation
  878. Noisy objective functions based on the f-divergence
  879. Non-blind image deconvolution using deep dual-pathway rectifier neural network
  880. Non-convex consensus ADMM for satellite precoder design
  881. Non-convex shredded signal reconstruction via sparsity enhancement
  882. Non-invasive gearbox fault diagnosis using scattering transform of acoustic emission
  883. Non-iterative impulse response shortening method for system latency reduction
  884. Non-negative matrix factorization of signals with overlapping events for event detection applications
  885. Non-orthogonal constrained independent vector analysis: Application to data fusion
  886. Non-parallel voice conversion using i-vector PLDA: towards unifying speaker verification and transformation
  887. Non-parametric analog Joint Source Channel Coding for amplify-and-forward two-hop networks
  888. Non-parametric spectrum cartography using adaptive radial basis functions
  889. Non-separable quadruple lifting structure for four-dimensional integer Wavelet Transform with reduced rounding noise
  890. Noncontact respiration monitoring of multiple closely positioned patients using ultra-wideband array radar with adaptive beamforming technique
  891. Nonparametric learning for Hidden Markov Models with preferential attachment dynamics
  892. Normal-to-shouted speech spectral mapping for speaker recognition under vocal effort mismatch
  893. Novel Amplitude Scaling method for bilinear frequency Warping-based Voice Conversion
  894. Novel medical video compression methods over lossless HEVC coder
  895. Novelty detection for predicting falls risk using smartphone gait data
  896. Null-steering beamformer for acoustic feedback cancellation in a multi-microphone earpiece optimizing the maximum stable gain
  897. Numerical filtering of linear state-space models with Markov switching
  898. ORGB: Offset correction in RGB color space for illumination-robust image processing
  899. Object detection refinement using Markov random field based pruning and learning based rescoring
  900. Objective assessment of pathological speech using distribution regression
  901. Objective characterization of audio signal quality: Applications to music collection description
  902. Omnidirectional bats, point-to-plane distances, and the price of uniqueness
  903. On DNN posterior probability combination in multi-stream speech recognition for reverberant environments
  904. On TOA estimation of vibration signals for localizing impacts on solid surfaces
  905. On classification of distorted images with deep convolutional neural networks
  906. On classification of environmental acoustic data using crowds
  907. On cognitive radio systems with directional antennas and imperfect spectrum sensing
  908. On methods for privacy-preserving energy disaggregation
  909. On mitigation of pilot spoofing attack
  910. On mutual coupling for ULAs: Estimating AoAs in the presence of more coupling parameters
  911. On random weights for texture generation in one layer CNNS
  912. On relationships between amplitude and phase of short-time Fourier transform
  913. On saturation of the Cramér Rao Bound for Sparse Bayesian Learning
  914. On spatial dependency in molecular distributed detection
  915. On spectrogram local maxima
  916. On the bias of pseudolinear estimators for time-of-arrival based localization
  917. On the impact of non-modal phonation on phonological features
  918. On the information rate of speech communication
  919. On the robustness of constrained convolutional neural networks to JPEG post-compression for image resampling detection
  920. On the role of head motion in affective expression
  921. On the security of block scrambling-based ETC systems against jigsaw puzzle solver attacks
  922. On time-frequency mask estimation for MVDR beamforming with application in robust speech recognition
  923. One-bit sparse array DOA estimation
  924. Online Empirical Mode Decomposition
  925. Online action detection and forecast via Multitask deep Recurrent Neural Networks
  926. Online environmental adaptation of CNN-based acoustic models using spatial diffuseness features
  927. Online learning of time-frequency patterns
  928. Online secondary path modelling in wave-domain active noise control
  929. Optical Tomography based on a nonlinear model that handles multiple scattering
  930. Optical-flow features empirical mode decomposition for motion anomaly detection
  931. Optimal achievable rate trade-off in cooperative cognitive radio systems
  932. Optimal biased estimation using Lehmann-unbiasedness
  933. Optimal low-rank Dynamic Mode Decomposition
  934. Optimal sparse L1-norm principal-component analysis
  935. Optimal transmit strategy for MIMO channels with joint sum and per-antenna power constraints
  936. Optimization of compound regularization parameters based on Stein's unbiased risk estimate
  937. Optimization over directed graphs: Linear convergence rate
  938. Optimized compressive sensing-based direction-of-arrival estimation in massive MIMO
  939. Optimizing Non Constant Luminance into Constant Luminance for High Dynamic Range Video Distribution
  940. Optimizing neural-network supported acoustic beamforming by algorithmic differentiation
  941. Optimizing speaker-specific filter banks for speaker verification
  942. Optimum array configurations of maximum output SNR for quiescent beamforming
  943. Orthogonal precoding for sidelobe suppression in DFT-based systems using block reflectors
  944. Overlapping sound event detection with supervised Nonnegative Matrix Factorization
  945. P-leader multifractal analysis for text type identification
  946. POKEMON: A non-linear beamforming algorithm for 1-bit massive MIMO
  947. PPG-based heart rate estimation using Wiener filter, phase vocoder and Viterbi decoding
  948. Pairwise learning using multi-lingual bottleneck features for low-resource query-by-example spoken term detection
  949. Parallel phonetically aware DNNs and LSTM-RNNS for frame-by-frame discriminative modeling of spoken language identification
  950. Parallelized Stochastic Gradient Markov Chain Monte Carlo algorithms for non-negative matrix factorization
  951. Parameter-free Plug-and-Play ADMM for image restoration
  952. Parameter-free automated extraction of neuronal signals from calcium imaging data
  953. Parametric estimation of spectrum driven by an exogenous signal
  954. Parametrized design of the generalized sequential probability ratio test
  955. Parsimonious Online Learning with Kernels via sparse projections in function space
  956. Part-level fully convolutional networks for pedestrian detection
  957. Partial image blur detection and segmentation from a single snapshot
  958. Particle PHD filter based multi-target tracking using discriminative group-structured dictionary learning
  959. Particle flow SMC delta-GLMB filter
  960. Particle flow for sequential Monte Carlo implementation of probability hypothesis density
  961. Partitioned Hierarchical alternating least squares algorithm for CP tensor decomposition
  962. Partitioned inverse image reconstruction for millimeter-wave SAR imaging
  963. Patch-based multiple view image denoising with occlusion handling
  964. Patch-based segmentation of overlapping cervical cells using active contour with local edge information
  965. Pattern recognition of functional brain networks
  966. Peak load minimization in load coupled interference networks
  967. Penalty dual decomposition method with application in signal processing
  968. Perceptual evaluation of a multiband acoustic crosstalk canceler using a linear loudspeaker array
  969. Performance analysis for time-of-arrival estimation with oversampled low-complexity 1-bit a/d conversion
  970. Performance analysis of (TDD) massive MIMO with Kalman channel prediction
  971. Performance analysis of an AoA estimator in the presence of more mutual coupling parameters
  972. Performance analysis of coarray-based MUSIC and the Cramér-Rao bound
  973. Performance bounds for Poisson compressed sensing using Variance Stabilization Transforms
  974. Performance of time delay estimation in a cognitive radar
  975. Performance trade-off in an adaptive IEEE 802.11AD waveform design for a joint automotive radar and communication system
  976. Permutation invariant training of deep models for speaker-independent multi-talker speech separation
  977. Personalized acoustic modeling by weakly supervised multi-task deep learning using acoustic tokens discovered from unlabeled data
  978. Personalized video emotion tagging through a topic model
  979. Personalized video preference estimation based on early fusion using multiple users' viewing behavior
  980. Perturbation analysis of Joint Eigenvalue Decomposition Algorithms
  981. Phase Congruency for image understanding with applications in computational seismic interpretation
  982. Phase estimation in single-channel speech enhancement using phase invariance constraints
  983. Phase reconstruction method based on time-frequency domain harmonic structure for speech enhancement
  984. Phase retrieval from STFT measurements via non-convex optimization
  985. Phase retrieval with a multivariate Von Mises prior: From a Bayesian formulation to a lifting solution
  986. Phase unmixing: Multichannel source separation with magnitude constraints
  987. Phase-dependent anisotropic Gaussian model for audio source separation
  988. Phaseless super-resolution in the continuous domain
  989. Phonological content impact on wrongful convictions in Forensic Voice Comparison context
  990. Pickup position and plucking point estimation on an electric guitar
  991. Pilot precoding and combining in multiuser MIMO networks
  992. Pitch contour tracking in music using Harmonic Locked Loops
  993. Pitch-based non-intrusive objective intelligibility prediction
  994. Polarimetric radar crosstalk removal during sparse image formation
  995. Polarization spectrogram of bivariate signals
  996. Polyphonic piano note transcription with non-negative matrix factorization of differential spectrogram
  997. Portable modeling of virtual physics for audio and haptic interaction design
  998. Pose-based composition improvement for portrait photographs
  999. Post-ICA phase de-noising for resting-state complex-valued FMRI data
  1000. Power-law stochastic neighbor embedding

Looking for submission deadlines instead? See the conference deadline calendar.