← All conferences

ICASSP 2015 Accepted Papers

The full list of 1,198 papers accepted at ICASSP 2015 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

  1. 3D SAR beamforming under a foliage canopy from a single pass
  2. 3D numerical modeling of parametric speaker using finite-difference time-domain
  3. 4D model-based iterative reconstruction from interlaced views
  4. <30 mW rectangular-to-polar conversion processor in 802.11ad polar transmitter
  5. A 3D model for room boundary estimation
  6. A Bayesian approach for the joint estimation of the multifractality parameter and integral scale based on the Whittle approximation
  7. A Bayesian approach to spatial filtering and diffuse power estimation for joint dereverberation and noise reduction
  8. A Bernoulli filter approach to detection and estimation of hidden Markov models using cluttered observation sequences
  9. A CFAR algorithm based on summations processing
  10. A Conditional Random Field system for beat tracking
  11. A Dimensional Contextual Semantic Model for music description and retrieval
  12. A Gaussian Mixture Model layer jointly optimized with discriminative features within a Deep Neural Network architecture
  13. A ROI-based self-embedding method with high recovery capability
  14. A beamformed alamouti amplify-and-forward scheme in multigroup multicast cloud-relay networks
  15. A bio-inspired logical process for saliency detections in cognitive crowd monitoring
  16. A cluster-voting approach for speaker diarization and linking of Australian broadcast news recordings
  17. A compact representation of sensor fingerprint for camera identification and fingerprint matching
  18. A comparative study of spectral clustering for i-vector-based speaker clustering under noisy conditions
  19. A comparison of extreme learning machines and back-propagation trained feed-forward networks processing the mnist database
  20. A consensus-based decentralized algorithm for non-convex optimization with application to dictionary learning
  21. A constrained hybrid Cramér-Rao bound for parameter estimation
  22. A cooperative jamming protocol for physical layer security in wireless networks
  23. A correctness result for online robust PCA
  24. A crosslinguistic study of prosodic focus
  25. A crosstalk-based linear filter in biochemical signal transduction pathways for the internet of bio-things
  26. A data-driven approach for matching clinical expertise to individual cases
  27. A data-driven color feature learning scheme for image retrieval
  28. A decomposition method for optimal user assignment in cellular networks with orthogonal transmissions
  29. A deep neural network approach to speech bandwidth expansion
  30. A deep neural network for time-domain signal reconstruction
  31. A deep recurrent approach for acoustic-to-articulatory inversion
  32. A directional noise suppressor with a specified beamwidth
  33. A discriminative post-filter for speech enhancement in hearing aids
  34. A down-mixing method for 22.2 multichannel system reproduction
  35. A dynamic programming variant of non-negative matrix deconvolution for the transcription of struck string instruments
  36. A factorization network based method for multi-lingual domain classification
  37. A fast audio search method based on skipping irrelevant signals by similarity upper-bound calculation
  38. A fast hyperplane-based MVES algorithm for hyperspectral unmixing
  39. A feedback framework for improved chord recognition based on NMF-based approximate note transcription
  40. A gradient adaptive population importance sampler
  41. A graph Laplacian regularization for hyperspectral data unmixing
  42. A hearing model to estimate mandarin speech intelligibility for the hearing impaired patients
  43. A histogram density modeling approach to music emotion recognition
  44. A hybrid GMM/SMC diffusion Bernoulli Filter for joint distributed detection and tracking
  45. A hybrid edge-preserving image smoothing scheme for noise removal
  46. A hybrid partial sum computation unit architecture for list decoders of polar codes
  47. A hybrid recurrent neural network for music transcription
  48. A hybrid speaker array-headphone system for immersive 3D audio reproduction
  49. A joint audio-visual approach to audio localization
  50. A keyword-aware grammar framework for LVCSR-based spoken keyword search
  51. A language space representation for speech recognition
  52. A language-based generative model framework for behavioral analysis of couples' therapy
  53. A learning-based approach to direction of arrival estimation in noisy and reverberant environments
  54. A low complexity iterative soft-decision feedback MMSE-PIC detection algorithm for massive MIMO
  55. A low complexity optimization algorithm for zero-forcing precoding under per-antenna power constraints
  56. A low-frequency superdirective acoustic vector sensor array
  57. A machine-hearing system exploiting head movements for binaural sound localisation in reverberant conditions
  58. A maximum correntropy criterion for robust multidimensional scaling
  59. A mixture of experts approach towards intelligibility classification of pathological speech
  60. A model of bottom-up visual attention using cortical magnification
  61. A mouth full of words: Visually consistent acoustic redubbing
  62. A multi-channel corpus for distant-speech interaction in presence of known interferences
  63. A multi-level representation of f0 using the continuous wavelet transform and the Discrete Cosine Transform
  64. A multi-modal approach using a non-parametric model to extract fetal ECG
  65. A multi-slice model observer for medical image quality assessment
  66. A multiple covariance approach for cell detection of Gram-stained smears images
  67. A new ADMM algorithm for the Euclidean Median and its application to robust patch regression
  68. A new Bayesian unmixing algorithm for hyperspectral images mitigating endmember variability
  69. A new alpha and gamma based mixture approximation for heavy-tailed Rayleigh distribution
  70. A new framework for solving dynamic scheduling games
  71. A new robust and efficient estimator for ill-conditioned linear inverse problems with outliers
  72. A new study of GMM-SVM system for text-dependent speaker recognition
  73. A nonmonotone learning rate strategy for SGD training of deep neural networks
  74. A novel QRS complex detection on ECG with motion artifact during exercise
  75. A novel Time-Delay-of-Arrival estimation technique for multi-microphone audio processing
  76. A novel approach for automatic acoustic novelty detection using a denoising autoencoder with bidirectional LSTM neural networks
  77. A novel filtering based approach for epoch extraction
  78. A novel image secret sharing scheme with meaningful shares
  79. A novel pooling strategy for Full Reference Image Quality Assessment based on harmonic means
  80. A novel ranking method for multiple classifier systems
  81. A novel sinusoidal approach to audio signal frame loss concealment and its application in the new evs codec standard
  82. A novel static parameter calculation method for model compensation
  83. A numerical implementation of gridless compressed sensing
  84. A pairwise algorithm for pitch estimation and speech separation using deep stacking network
  85. A parametric Bayesian RMC gamma-ray image reconstruction
  86. A parametric modeling approach for wireless capsule endoscopy hazy image restoration
  87. A performance study of the tangent distance method in transformation-invariant image classification
  88. A physiological correlate of electroacoustic pitch matching in cochlear implant users with residual hearing
  89. A priori SAP estimator based on the magnitude square coherence for dual-channel microphone system
  90. A probabilistic interpretation of sampling theory of graph signals
  91. A probabilistic least-mean-squares filter
  92. A proof of Hirschman Uncertainty invariance to the order of Rényi entropy for Picket Fence signals, and its relevance in a simplistic recognition experiment
  93. A proximal gradient algorithm for decentralized nondifferentiable optimization
  94. A quaternion frequency estimator for three-phase power systems
  95. A random block-coordinate primal-dual proximal algorithm with application to 3D mesh denoising
  96. A randomized dual consensus ADMM method for multi-agent distributed optimization
  97. A robust motion detection algorithm on noisy videos
  98. A robust online subspace estimation and tracking algorithm
  99. A robust region-based near-field beamformer
  100. A robust sparse approach to acoustic impulse response shaping
  101. A sequential dictionary learning algorithm with enforced sparsity
  102. A signal processing scheme for a multichannel passive radar system
  103. A simple method for DOA estimation in the presence of unknown nonuniform noise
  104. A simple modification to facilitate robust generalized sidelobe canceller for hearing aids
  105. A simple user interface system for recovering patterns repeating in time and frequency in mixtures of sounds
  106. A stackelberg game-based energy trading scheme for power beacon-assisted wireless-powered communication
  107. A state-space approach for the analysis of wave and diffusion fields
  108. A state-space partitioned-block adaptive filter for echo cancellation using inter-band correlations in the Kalman gain computation
  109. A statistical comparison between music and G-music
  110. A stochastic behavior analysis of stochastic restricted-gradient descent algorithm in reproducing kernel hilbert spaces
  111. A study on joint beamforming and spectral enhancement for robust speech recognition in reverberant environments
  112. A temporal limits encoder for cochlear implants
  113. A tensor LMS algorithm
  114. A tensor-based subspace wall clutter mitigation method for through-the-wall radar imaging
  115. A three-stage framework to active source localization from a binaural head
  116. A two channel approach for system approximation with general measurement functionals
  117. A unified approach for hybrid source localization based on ranges and video
  118. A unified framework for filterbank and time-frequency basis vectors in ASR frontends
  119. A unified probabilistic framework for robust decoding of linear barcodes
  120. A virtual bass system with improved overflow control
  121. A virtual resampling technique for algebraic two-dimensional phase unwrapping
  122. A workload balanced parallel view synthesis for FTV
  123. AA spectral space warping approach to cross-lingual voice transformation in HMM-based TTS
  124. ASIC implementation of a computationally efficient compressive sensing detection method using least squares optimization in 45 nm CMOS technology
  125. ASR error detection and recognition rate estimation using deep bidirectional recurrent neural networks
  126. Accelerating and deceleratingmin-sum-based gear-shift LDPC decoders
  127. Accurate analysis method of background ionosphere effects on Geosynchronous SAR focusing
  128. Accurate kernel-based spectrum sensing for Gaussian and non-Gaussian noise models
  129. Achievable degrees-of-freedom of (n, K)-user interference channel with distributed beamforming
  130. Achieving high resolution for super-resolution via reweighted atomic norm minimization
  131. Acoustic and para-verbal indicators of persuasiveness in social multimedia
  132. Acoustic event source localization for surveillance in reverberant environments supported by an event onset detection
  133. Acoustic feature extraction by tensor-based sparse representation for sound effects classification
  134. Acoustic scene analysis from acoustic event sequence with intermittent missing event
  135. Active feedback noise control in the presence of impulsive disturbances
  136. Active learning of self-concordant like multi-index functions
  137. Active matching for patch adaptivity in nonlocal means image denoising
  138. Activity-mapping non-negative matrix factorization for exemplar-based voice conversion
  139. Adaptive Bayesian tracking with unknown time-varying sensor network performance
  140. Adaptive Sensor Data Compression in IoT systems: Sensor data analytics based approach
  141. Adaptive censoring for large-scale regressions
  142. Adaptive damping and mean removal for the generalized approximate message passing algorithm
  143. Adaptive differential microphone arrays used as a front-end for an automatic speech recognition system
  144. Adaptive multicast beamforming: Guaranteed convergence and state-of-art performance at low complexity
  145. Adaptive neural matching online spike sorting VLSI chip design for wireless BCI implants
  146. Adaptive sensing resource allocation over multiple hypothesis tests
  147. Adaptive signal and system approximation and strong divergence
  148. Adaptive statistical utterance phonetization for French
  149. Additive noise compensation in the i-vector space for speaker recognition
  150. Advances in deep neural network approaches to speaker recognition
  151. Advances in low bitrate time-frequency coding
  152. Advantages of dynamic analysis in HOG-PCA feature space for video moving object classification
  153. Affective structure modeling of speech using probabilistic context free grammar for emotion recognition
  154. Algorithms and performance analysis for estimation of low-rank matrices with Kronecker structured singular vectors
  155. Aligning training modelswith smartphone properties in WiFi fingerprinting based indoor localization
  156. Alignment with intra-class structure can improve classification
  157. Alternating diffusion for common manifold learning with application to sleep stage assessment
  158. An Encryption-then-Compression system for JPEG 2000 standard
  159. An HMM-based formalism for automatic subword unit derivation and pronunciation generation
  160. An adaptive ECC scheme for dynamic protection of NAND Flash memories
  161. An adaptive low-complexity detection method for statistical signal transmission under time-varying channels
  162. An algorithm for the parameter estimation of multiple superimposed exponentials in noise
  163. An analysis of convolutional neural networks for speech recognition
  164. An approximate Newton method for distributed optimization
  165. An asymptotic LMPI test for cyclostationarity detection with application to cognitive radio
  166. An attack on antenna subset modulation for millimeter wave communication
  167. An effective key generation system using improved channel reciprocity
  168. An efficient interpolation filter VLSI architecture for HEVC
  169. An efficient kernel normalized least mean square algorithm with compactly supported kernel
  170. An energy-efficient memory-based high-throughput VLSI architecture for convolutional networks
  171. An evaluation of methodologies for melodic similarity in audio recordings of Indian art music
  172. An improved cross-correlation approach to parameter estimation based on fractional Fourier transform for ISAR motion compensation
  173. An improved variable step-size zero-point attracting projection algorithm
  174. An information-theoretic framework for automated discovery of prosodic cues to conversational structure
  175. An information-theoretic metric of fingerprint effectiveness
  176. An investigation into speaker informed DNN front-end for LVCSR
  177. An investigation of augmenting speaker representations to improve speaker normalisation for DNN-based speech recognition
  178. An iterative bayesian algorithm for block-sparse signal reconstruction
  179. An iterative deflation algorithm for exact CP tensor decomposition
  180. An iterative reweighted minimization framework for joint channel and power allocation in the OFDMA system
  181. An obfuscated radix-2 real FFT architecture
  182. An online EM algorithm in hidden (semi-)Markov models for audio segmentation and clustering
  183. An online algorithm for distributed dictionary learning
  184. An optimal dimensionality sampling scheme on the sphere for antipodal signals in diffusion magnetic resonance imaging
  185. An outreach after-school program to introduce high-school students to electrical engineering
  186. An overviewof the SKA project: Why take on this signal processing challenge?
  187. Analyses on empirical error minimization in multiple kernel regressors
  188. Analysis and automatic recognition of Human BeatBox sounds: A comparative study
  189. Analysis of H-sssi processes using the crossing tree: An alternative to wavelets
  190. Analysis of beamformer directed single-channel noise reduction system for hearing aid applications
  191. Analysis of singing voice for epoch extraction using Zero Frequency Filtering method
  192. Analysis of speech and language communication for cochlear implant users in noisy lombard conditions
  193. Analysis of target detection via matrix completion
  194. Anchor nodes refinement in joint localization and synchronization of a sensor node
  195. Annealed dropout trained maxout networks for improved LVCSR
  196. Annihilation-driven localised image edge models
  197. Annotating and categorizing competition in overlap speech
  198. Anti-cropping blind resynchronization for 3D watermarking
  199. Aperiodic waveforms with mismatched filtering for target detection in heavy clutter. Part I - SIMO radar architecture
  200. Aperiodicwaveforms with mismatched filtering for target detection in heavy clutter. part II - MIMO radar architecture
  201. Approach to frame-misalignment in physical-layer network coding
  202. Approximate infinite-dimensional Region Covariance Descriptors for image classification
  203. Arithmetic coding of speech and audio spectra using tcx based on linear predictive spectral envelopes
  204. Assessing range accuracy for bearings-only geolocation using optimal logarithmic spiral sensor path trajectories
  205. Assistive listening headsets for high noise environments: Protection and communication
  206. Asymptotic analysis of linear spectral statistics of the sample coherence matrix
  207. Asymptotic justification of bandlimited interpolation of graph signals for semi-supervised learning
  208. Asymptotic performance of the Low Rank Adaptive Normalized Matched Filter in a large dimensional regime
  209. Asymptotic properties of the robust ANMF
  210. Atom decomposition-based intonation modelling
  211. Attributing modelling errors in HMM synthesis by stepping gradually from natural to modelled speech
  212. Audio modeling and loudness estimation with IJDSP mobile simulations
  213. Audio source localization by optimal control of a mobile robot
  214. Audio source separation using a redundant library of source spectral bases for non-negative tensor factorization
  215. Audio synchronisation with a tunnel matrix for time series and dynamic programming
  216. Audiovisual speaker diarization of TV series
  217. Augmented covariance estimation with a cyclic approach in DOA
  218. Automated tracking of cells from phase contrast images by multiple hypothesis Kalman filters
  219. Automatic assessment of English learner pronunciation using discriminative classifiers
  220. Automatic broadcast news summarization via rank classifiers and crowdsourced annotation
  221. Automatic detection of voice onset time in dysarthric speech
  222. Automatic gain control and multi-style training for robust small-footprint keyword spotting with deep neural networks
  223. Automatic pronunciation verification for speech recognition
  224. Automatic target recognition using discrimination based on optimal transport
  225. Average recovery performances of non-perfectly informed compressed sensing: With applications to multiclass encryption
  226. Averaging based distributed estimation algorithm for sensor networks with quantized and directed communication
  227. Averaging random projection: A fast online solution for large-scale constrained stochastic optimization
  228. Base Station clustering in heterogeneous network with finite backhaul capacity
  229. Bayesian compressive sensing for DOA estimation using the difference coarray
  230. Bayesian narrowband interference mitigation in SC-FDMA
  231. Bayesian parameter estimation of Jump-Langevin systems for trend following in finance
  232. Bayesian path estimation using the spatial attributes of a road network
  233. Bayesian social learning in linear networks of agents with random behavior
  234. Bi-alternating direction method of multipliers over graphs
  235. Bi-directional differential beamforming for multi-antenna relaying
  236. Bidirectional recurrent neural network language models for automatic speech recognition
  237. Binaural localization of speech sources in 3-D using a composite feature vector of the HRTF
  238. Binaural multichannel Wiener filter with directional interference rejection
  239. Binaural speech enhancement with instantaneous coherence smoothing using the cepstral correlation coefficient
  240. Binomial classification based on DLENE features in sparse representation: Application in kidney detection in 3D ultrasound
  241. Bird-phrase segmentation and verification: A noise-robust template-based approach
  242. Blind bleed-through removal for scanned historical document images with conditional random fields
  243. Blind equalization and Automatic Modulation Classification based on pdf fitting
  244. Blind estimation of effective downlink channel gains in massive MIMO
  245. Blind signal separation of rational functions using Löwner-based tensorization
  246. Blind stain decomposition for histo-pathology images using circular nature of chroma components
  247. Blur kernel estimation approach to blind reverberation time estimation
  248. Body-structure based feature representation for person re-identification
  249. Buffer merging technique for minimizing memory footprints of Synchronous Dataflow specifications
  250. Building context-dependent DNN acoustic models using Kullback-Leibler divergence-based state tying
  251. C. elegans cell matching and tracking in a 4D imageing system
  252. Cell phone verification from speech recordings using sparse representation
  253. Cepstral noise subtraction for robust automatic speech recognition
  254. Challenges in deploying a microphone array to localize and separate sound sources in real auditory scenes
  255. Change detection for optical and radar images using a Bayesian nonparametric model coupled with a Markov random field
  256. Channel adaptation of plda for text-independent speaker verification
  257. Chasing butterflies: In search of efficient dictionaries
  258. Classification of whale vocalizations using the Weyl transform
  259. Classifying phonological categories in imagined and articulated speech
  260. Closed-form Cramer-Rao lower bounds for DOA estimation from turbo-coded square-QAM-modulated transmissions
  261. Closed-form solution to directly design face waveforms for beampatterns using planar array
  262. Cluster adaptive training for deep neural network
  263. CoCE-SMART: Consensus clustering based on enhanced splitting-merging awareness tactics
  264. Coalitional game theoretic approach to distributed adaptive parameter estimation
  265. Coded aperture compressive 3-D LIDAR
  266. Cognitive biases in Bayesian updating and optimal information sequencing
  267. Coherent channel based subband multichannel dereverberation
  268. Coherent modification of pitch and energy for expressive prosody implantation
  269. Collaborative compressive X-ray image reconstruction
  270. Collaborative filtering based on group coordinates for smoothing and directional sharpening
  271. Collaborative randomized beamforming for phased array radio interferometers
  272. Color description of low resolution images using fast bitwise quantization and border-interior classification
  273. Color facial expression recognition based on color local features
  274. Combination of search techniques for improved spotting of OOV keywords
  275. Combination of two-dimensional cochleogram and spectrogram features for deep learning-based ASR
  276. Combined estimation of camera link models for human tracking across nonoverlapping cameras
  277. Combining Compressed Sensing with motion correction in acquisition and reconstruction for PET/MR
  278. Combining SGMM speaker vectors and KL-HMM approach for speaker diarization
  279. Combining information display and visible light wireless communication
  280. Combining robust spike coding with spiking neural networks for sound event classification
  281. Combining sparse NMF with deep neural network: A new classification-based approach for speech enhancement
  282. Combining sparsity with rank-deficiency for energy efficient EEG sensing and transmission over Wireless Body Area Network
  283. Combining two phase codes to extend the radar unambiguous range and get a trade-off in terms of performance for any clutter
  284. Common components analysis via linked blind source separation
  285. Common part estimation of acoustic feedback paths in hearing aids optimizing maximum stable gain
  286. Comparative performance evaluation of error regularized Turbo-MIMO MMSE-SIC detectors in Gaussian channels
  287. Compensating for asynchronies between musical voices in score-performance alignment
  288. Compressed sensing based multi-user millimeter wave systems: How many measurements are needed?
  289. Compressed sensing joint range and cross-range MIMO Radar imaging
  290. Compressive graph clustering from random sketches
  291. Compressive parameter estimation via approximate message passing
  292. Computationally deconstructing movie narratives: An informatics approach
  293. Computationally efficient radio astronomical image formation using constrained least squares and and the MVDR beamformer
  294. Computing multistatic passive radar CFAR thresholds from surveillance-only data
  295. Consensus for the distributed estimation of point diffusion sources in sensor networks
  296. Consistency of ℓ1-regularized maximum-likelihood for compressive Poisson regression
  297. Constrained state estimation in particle filters
  298. Constructing long short-term memory based deep recurrent neural networks for large vocabulary speech recognition
  299. Content-based recommendations with approximate integer division
  300. Content-based recommender systems for spoken documents
  301. Context adaptive deep neural networks for fast acoustic model adaptation
  302. Context dependent phone models for LSTM RNN acoustic modelling
  303. Contextual spoken language understanding using recurrent neural networks
  304. Continuous visual speech recognition for audio speech enhancement
  305. Convergence analysis of alternating direction method of multipliers for a family of nonconvex problems
  306. Convergence analysis of the augmented complex klms algorithm with pre-tuned dictionary
  307. Convergence of an inertial proximal method for l1-regularized least-squares
  308. Convolutional Neural Networks-based continuous speech recognition using raw speech signal
  309. Convolutional, Long Short-Term Memory, fully connected Deep Neural Networks
  310. Cooperative localization with information-seeking control
  311. Cooperative self-localization in asynchronous sensors networks based on TOA from transmitters at unknown locations
  312. Copingwith channel mismatch in Query-by-Example - But QUESST 2014
  313. Coprime DFT filter bank design: Theoretical bounds and guarantees
  314. Coprime arrays and samplers for space-time adaptive processing
  315. Copy-move detection of audio recording with pitch similarity
  316. Correlation-aware sparsity-enforcing sensor placement for spatio-temporal field estimation
  317. Cost-sensitive ensemble classifiers for microwave breast cancer detection
  318. Coupled fisher discrimination dictionary learning for single image super-resolution
  319. Coupled learning based on singular-values-unique and hog for face hallucination
  320. Covariance tracking from sketches of rapid data streams
  321. Cramér-Rao-type bound for state estimation in linear discrete-time system with unknown system parameters
  322. Critically sampled graph wavelets converted from linear-phase biorthogonal wavelets
  323. Cross-corpus depression prediction from speech
  324. Cross-domain cooperative deep stacking network for speech separation
  325. Cross-lingual lexical language discovery from audio data using multiple translations
  326. Curvature-based optimization of the trade-off parameter in the speech distortion weighted multichannel wiener filter
  327. Cyber-physical systems: Dynamic sensor attacks and strong observability
  328. Data augmentation for deep convolutional neural network acoustic modeling
  329. Deconvolution using the adaptive selective sidelobe canceller beamformer
  330. Deep NMF for speech separation
  331. Deep autoencoders augmented with phone-class feature for reverberant speech recognition
  332. Deep convolutional activation features for large scale Brain Tumor histopathology image classification and segmentation
  333. Deep convolutional neural networks for acoustic modeling in low resource languages
  334. Deep multimodal learning for Audio-Visual Speech Recognition
  335. Deep neural network based instrument extraction from music
  336. Deep neural networks employing Multi-Task Learning and stacked bottleneck features for speech synthesis
  337. Deep neural networks for cochannel speaker identification
  338. Deep neural networks for estimating speech model activations
  339. Deep neural support vector machines for speech recognition
  340. Deep recurrent regularization neural network for speech recognition
  341. Deformable multiple-kernel based human tracking using a moving camera
  342. Delay control for CDF scheduling with deadlines
  343. Delayless speech enhancement with a virtual zero-phase response using a prediction of periodic signal components
  344. Demixing multivariate-operator self-similar processes
  345. Dense and continuous depth estimation using a sliding camera
  346. Dense correspondence based prediction for image set compression
  347. Density estimation by entropy maximization with kernels
  348. Depth image super-resolution using internal and external information
  349. Dereverberation sweet spot dilation with combined channel equalization and beamforming
  350. Design and analysis of miniature and three tiered B-format microphones manufactured using 3D printing
  351. Design of signal-matched critically sampled FIR rational filterbank
  352. Designing multichannel source separation based on single-channel source separation
  353. Destination inference using bridging distributions
  354. Detecting hidden cliques from noisy observations
  355. Detecting kangaroos in the wild: the first step towards automated animal surveillance
  356. Detecting laterality and nasality in speech with the use of a multi-channel recorder
  357. Detecting rare events using Kullback-Leibler divergence
  358. Detecting semantic concepts in consumer videos using audio
  359. Detection Aided Multistatic Velocity Backprojection for passive radar
  360. Detection and recognition of deformable objects using structured dimensionality reduction
  361. Detection and suppression of keyboard transient noise in audio streams with auxiliary keybed microphone
  362. Detection of depression in adolescents based on statistical modeling of emotional influences in parent-adolescent conversations
  363. Detection of pilot spoofing attack in multi-antenna systems via energy-ratio comparison
  364. Determining the number of correlated signals between two data sets using PCA-CCA when sample support is extremely small
  365. Deterministic constructions of binary measurement matrices with various sizes
  366. Diarization resegmentation in the factor analysis subspace
  367. Dictionary-based online kernel principal subspace analysis with double orthogonality preservation
  368. Differentiable pooling for unsupervised speaker adaptation
  369. Diffusion filtration with approximate Bayesian computation
  370. Direct-ambient decomposition using parametric wiener filtering with spatial cue control
  371. Direct-to-Reverberant Ratio estimation using a null-steered beamformer
  372. Direction-finding based on the theory of super-resolution in sparse recovery algorithms
  373. Direction-of-arrival and diffuseness estimation above spatial aliasing for symmetrical directional microphone arrays
  374. Direction-of-arrival estimation of speech sources under aliasing conditions
  375. Directional bilateral filters
  376. Directionality assessment of adaptive binaural beamforming with noise suppression in hearing aids
  377. Directly modeling speech waveforms by neural networks for statistical parametric speech synthesis
  378. Discriminative method for recurrent neural network language models
  379. Discriminative spectral learning of hidden markov models for human activity recognition
  380. Discriminative uncertainty estimation for noise robust ASR
  381. Disparity-compensated total-variation minimization for compressed-sensed multiview image reconstruction
  382. Distributed AOA-based source positioning in NLOS with sensor networks
  383. Distributed Kalman Filtering with quantized sensing state
  384. Distributed algorithm for graph signal inpainting
  385. Distributed beamforming for cooperative networks with widely-linear processing at the relays and the receiver
  386. Distributed black-box optimization of nonconvex functions
  387. Distributed dense stereo matching for 3D reconstruction using parallel-based processing advantages
  388. Distributed dialogue policies for multi-domain statistical dialogue management
  389. Distributed kernel learning using Kernel Recursive Least Squares
  390. Distributed primal strategies outperform primal-dual strategies over adaptive networks
  391. Distributed robust change point detection for autoregressive processes with an application to distributed voice activity detection
  392. Distributed robust labeling of audio sources in heterogeneous wireless sensor networks
  393. Distributed signal estimation in a wireless sensor network with partially-overlapping node-specific interests or source observability
  394. Distributed spatio-temporal multi-target association and tracking
  395. Distributed target tracking under communication constraints
  396. Distributed tls estimation under random data faults
  397. Distributed topology identification for point process dynamic networks
  398. Distributions of projections of uniformly distributed K-frames
  399. Diversity combining in wireless relay networks with partial channel state information
  400. Doa estimation by covariance matrix sparse reconstruction of coprime array
  401. Doa estimation of nonparametric spreading spatial spectrum based on bayesian compressive sensing exploiting intra-task dependency
  402. Document-specific context plsa language model for speech recognition
  403. Dominant SIFT: A novel compact descriptor
  404. Double differential transmission for two-way relay systems with unknown carrier frequency offsets
  405. Double-layer neighborhood graph based similarity search for fast query-by-example spoken term detection
  406. Double-talk detection in acoustic echo cancellers using zero-crossings rate
  407. Downbeat tracking with multiple features and deep neural networks
  408. Downsampling for sparse subspace clustering
  409. Dual-exposure image registration for HDR processing
  410. Dynamic ROI based on K-means for remote photoplethysmography
  411. Dynamic sparse state estimation using ℓ1-ℓ1 minimization: Adaptive-rate measurement bounds, algorithms and applications
  412. Dynamic zero-point attracting projection for time-varying sparse signal recovery
  413. EEG dimensionality reduction in automatic identification of synonymy
  414. EEG signal enhancement using multi-channel wiener filter with a spatial correlation prior
  415. EEG source reconstruction performance as a function of skull conductance contrast
  416. Effects of feature type, learning algorithm and speaking style for depression detection from speech
  417. Efficient FFT method for modelling performance of radars with scan-to-scan feedback integration
  418. Efficient FxLMS algorithm with simplified secondary path models
  419. Efficient and accurate multivariate class conditional densities using copula
  420. Efficient audio declipping using regularized least squares
  421. Efficient blind estimation of subband reverberation time from speech in non-diffuse environments
  422. Efficient coding strategy for HEVC performance improvement by exploiting motion features
  423. Efficient collaborative sparse channel estimation in massive MIMO
  424. Efficient construction of dictionaries for kernel adaptive filtering in a dynamic environment
  425. Efficient detection and localization on graph structured data
  426. Efficient filtering and sampling for a class of time-varying linear systems
  427. Efficient handling of mode switching and speech transitions in the EVS codec
  428. Efficient image categorization with sparse Fisher vector
  429. Efficient linear combination of partial Monte Carlo estimators
  430. Efficient manifold preserving audio source separation using locality sensitive hashing
  431. Efficient model choice and parameter estimation by using nested sampling applied in Eddy-Current Testing
  432. Efficient multichannel nonnegative matrix factorization exploiting rank-1 spatial model
  433. Efficient spectrogram-based binary image feature for audio copy detection
  434. Efficient update of persistent particles in the SMC-PHD filter
  435. Embedded real-time localization of UAV based on an hybrid device
  436. Emotion recognition using synthetic speech as neutral reference
  437. Employment of Subspace Gaussian Mixture Models in speaker recognition
  438. Energy-efficient precoding matrix design for relay-aided multiuser downlink networks
  439. Enhanced noninvasive imaging system for dispersive highly coherent space
  440. Enhanced robot audition by dynamic acoustic sensing in moving humanoids
  441. Enhanced time domain packet loss concealment in switched speech/audio codec
  442. Enhancing LDPC code performance using pilot bits
  443. Enhancing automatically discovered multi-level acoustic patterns considering context consistency with applications in spoken term detection
  444. Enhancing class discrimination in Kernel Discriminant Analysis
  445. Enhancing local - Transmitting less - Improving global
  446. Enhancing sparse voice annotation for semantic retrieval of personal photos by continuous space word representations
  447. Entropy analysis of i-vector feature spaces in duration-sensitive speaker recognition
  448. Error diffused intra prediction for HEVC
  449. Esprit-type algorithms for a received mixture of circular and strictly non-circular signals
  450. Estimate articulatory MRI series from acoustic signal using deep architecture
  451. Estimating confidence scores on ASR results using recurrent neural networks
  452. Estimating double thumbnails for music recordings
  453. Estimating link-dependent Origin-Destination matrices from sample trajectories and traffic counts
  454. Estimation of multipath propagation delays and interaural time differences from 3-D head scans
  455. Estimation of rapidly varying sea clutter using nearest Kronecker product approximation
  456. Estimation of relative transfer function in the presence of stationary noise based on segmental power spectral density matrix subtraction
  457. Estimation of the invariant and variant characteristics in speech articulation and its application to speaker identification
  458. Evaluating Deep Scattering Spectra with deep neural networks on large scale spontaneous speech task
  459. Evaluation of linear regression for speaker adaptation in HMM-based articulatory movements estimation
  460. Evaluation of speech inverse filtering techniques using a physiologically based synthesizer
  461. Exact asymptotics of distributed detection over adaptive networks
  462. Exemplar-based large vocabulary speech recognition using k-nearest neighbors
  463. Exemplar-based speech enhancement for deep neural network based automatic speech recognition
  464. Explicit order model for region-based level set segmentation
  465. Explicit versus implicit source estimation for blind multiple input single output system identification
  466. Exploiting FRI signal structure for sub-Nyquist sampling and processing in medical ultrasound
  467. Exploiting subclass information in one-class support vector machine for video summarization
  468. Exploring multi-channel features for denoising-autoencoder-based speech enhancement
  469. Extracting deep bottleneck features for visual speech recognition
  470. Extracting singing voice from music recordings by cascading audio decomposition techniques
  471. Extraction of pitch register from expressive speech in Japanese
  472. FRI sampling and reconstruction of asymmetric pulses
  473. Face detection using Local Hybrid Patterns
  474. Face hallucination via Cauchy regularized sparse representation
  475. Face recognition for great apes: Identification of primates in videos
  476. Face-based Active Authentication on mobile devices
  477. Factorization for analog-to-digital matrix multiplication
  478. Fairness considerations in full-duplex MIMO interference channels
  479. Far-field speech recognition using CNN-DNN-HMM with convolution in time
  480. Fast DNN training based on auxiliary function technique
  481. Fast and efficient intra coding techniques for smooth regions in screen content coding based on boundary prediction samples
  482. Fast and robust EM-based IRLS algorithm for sparse signal recovery from noisy measurements
  483. Fast approximate i-vector estimation using PCA
  484. Fast compressive phase retrieval from Fourier measurements
  485. Fast convex optimization for connectivity enforcement in gene regulatory network inference
  486. Fast efficient and scalable Core Consistency Diagnostic for the parafac decomposition for big sparse tensors
  487. Fast image interpolation with decision tree
  488. Fast implementation of a family of memory proportionate affine projection algorithm
  489. Fast line and circle detection using inverted gradient hash maps
  490. Fast magnetic susceptibility reconstruction using L0 norm of gradient
  491. Fast realistic refocusing for sparse light fields
  492. Feasibility of FRI-based square-wave reconstruction with quantization error and integrator noise
  493. Feature enhancement based on generative-discriminative hybrid approach with gmms and DNNS for noise robust speech recognition
  494. Feedback-based handwriting recognition from inertial sensor data for wearable devices
  495. Finding line spectral frequencies using the fast fourier transform
  496. Fix it where it fails: Pronunciation learning by mining error corrections from speech logs
  497. Fixed point optimization of deep convolutional neural networks for object recognition
  498. Flexible spectrum coding in the 3GPP EVS codec
  499. Foreground suppression for capturing and reproduction of crowded acoustic environments
  500. Forensic voice comparison with monophthongal formant trajectories - a likelihood ratio-based discrimination of "schwa" vowel acoustics in a close social group of young Australian females
  501. Forward stereo obstacle detection with Weighted Hough Transform and local temporal correlation
  502. Fractional spatial reuse precoding for MIMO downlink networks
  503. Free energy for speech recognition
  504. Frequency hopping waveforms for continuous active sonar
  505. Frequency-domain Comfort Noise Generation for Discontinuous Transmission in EVS
  506. From Simulink to smartphone: Signal processing application examples
  507. From antennas to multi-dimensional data cubes: The SKA data path
  508. Full-rank linear-chain NeuroCRF for sequence labeling
  509. Fusion of polarimetric radar images using hybrid matching pursuit
  510. Fusion of speaker and lexical information for topic segmentation: A co-segmentation approach
  511. Fusion of ultrasound harmonic imaging with clutter removal using sparse signal separation
  512. GLRT detection with unknown noise power in passive multistatic radar
  513. GNSS spoofing detection using multiple mobile COTS receivers
  514. GPU acceleration of Threat Map computation and application to selection of sonar field controls
  515. Gaussian signal detection by coprime sensor arrays
  516. General linear models under Rician noise for fMRI data
  517. General solution and approximate implementation of the multisensor multitarget CPHD filter
  518. Generalised array reconfiguration for adaptive beamforming by antenna selection
  519. Generalized Wiener filtering with fractional power spectrograms
  520. Generalized approximate message passing for cosparse analysis compressive sensing
  521. Generalized direct predistortion with adaptive crest factor reduction control
  522. Generative modeling of pseudo-target domain adaptation samples for whispered speech recognition
  523. Gesture recognition from magnetic field measurements using a bank of linear state space models and local likelihood filtering
  524. Globally optimized dynamic bit-allocation strategy for subband ADPCM-based low delay audio coding
  525. Gradient scan Gibbs sampler: An efficient high-dimensional sampler application in inverse problems
  526. Grapheme-to-phoneme conversion using Long Short-Term Memory recurrent neural networks
  527. Greedy minimization of l1-norm with high empirical success
  528. Ground moving target imaging by synthetic aperture radar based on an unified framework of keystone transformation
  529. HMM-based emphatic speech synthesis for corrective feedback in computer-aided pronunciation training
  530. HMM-based modelling of individual syllables for bird species recognition from audio field recordings
  531. Harmonic Vector Quantization
  532. Harmonic phase estimation in single-channel speech enhancement using von mises distribution and prior SNR
  533. Hierarchical Sparse and Collaborative Low-Rank representation for emotion recognition
  534. Hierarchical and Lossless Coding of audio objects in Dolby TrueHD
  535. High-resolution through-wall ghost imaging algorithm using chaotic modulated signal
  536. Higher-dimensional coherence of subspaces
  537. Histogram-PMHT with an evolving Poisson prior
  538. Horizontal flip-invariant sketch recognition via local patch hashing
  539. How to construct progressive visual cryptography schemes
  540. How to monitor and mitigate stair-casing in L1 trend filtering
  541. Hybrid digital and analog beamforming design for large-scale MIMO systems
  542. Hybrid multi-layer deep CNN/aggregator feature for image classification
  543. Hybrid vectorial and tensorial Compressive Sensing for hyperspectral imaging
  544. Hyper-spectral impulse denoising: A row-sparse Blind Compressed Sensing formulation
  545. I-vector based language modeling for query representation
  546. IVA algorithms using a multivariate Student's t source prior for speech source separation in real room environments
  547. Identification of the parametric array loudspeaker with a volterra filter using the sparse NLMS algorithm
  548. Identify Visual Human Signature in community via wearable camera
  549. Image colorization using hybrid domain transform
  550. Image compression via dense descriptors assisted synthesis
  551. Image interpolation using Gaussian Mixture Models with spatially constrained patch clustering
  552. Image masking schemes for local manifold learning methods
  553. Image matching for repetitive patterns
  554. Image quality assessment based on structure variance classification
  555. Image-guided customization of frequency-place mapping in cochlear implants
  556. Impact location estimation in anisotropic media
  557. Impact of station size on calibration of SKA-low
  558. Implementation of interconnective systems
  559. Implementation of the SVD-based precoding sub-system for the compressed beamforming weights feedback in IEEE 802.11n/ac WLAN
  560. Improved Parallel Feedback Active noise control using linear prediction for adaptive noise decomposition
  561. Improved direction finding using a maneuverable array of directional sensors
  562. Improved error resilience for volte and VoIP with 3GPP EVS channel aware coding
  563. Improved face-to-face communication using noise reduction and speech intelligibility enhancement
  564. Improved language identification using deep bottleneck network
  565. Improved linear least squares estimation using bounded data uncertainty
  566. Improved recognition of contact names in voice commands
  567. Improved speaker recognition using DCT coefficients as features
  568. Improved strategies for a zero oov rate LVCSR system
  569. Improved time-frequency trajectory excitation modeling for a statistical parametric speech synthesis system
  570. Improved view synthesis by motion warping and temporal hole filling
  571. Improvements on transducing syllable lattice to word lattice for keyword search
  572. Improvements to the IBM speech activity detection system for the DARPA RATS program
  573. Improving long short-term memory networks using maxout units for large vocabulary speech recognition
  574. Improving multiple-crowd-sourced transcriptions using a speech recogniser
  575. Improving music auto-tagging with trigger-based context model
  576. Improving n-gram probability estimates by compound-head clustering
  577. Improving out-domain PLDA speaker verification using unsupervised inter-dataset variability compensation approach
  578. Improving speech recognition in reverberation using a room-aware deep neural network and multi-task learning
  579. Improving the training and evaluation efficiency of recurrent neural network language models
  580. Impulsive noise detection in PLC with smoothed L0-norm
  581. Incorporating spatial information in binaural beamforming for noise suppression in hearing aids
  582. Individualizing a monaural beamformer for cochlear implant users
  583. Indoor mapping based on time delay estimation in wireless networks
  584. Indoor target tracking using high doppler resolution passive Wi-Fi radar
  585. Inferring causal connectivity in epileptogenic zone using directed information
  586. Influence of time-varying pitch on timbre: "Coherence and incoherence" based on spectral centroid
  587. Information extraction from large multi-layer social networks
  588. Informed monaural source separation of music based on convolutional sparse coding
  589. Informed source separation from monaural music with limited binary time-frequency annotation
  590. Integrated pronunciation learning for automatic speech recognition using probabilistic lexical modeling
  591. Integrating Gaussian mixtures into deep neural networks: Softmax layer with hidden variables
  592. Intelligibility evaluation of speech coding standards in severe background noise and packet loss conditions
  593. Interactive on-device Mobile Landmark Recognition with compact binary codes
  594. Interaural coherence preservation in MWF-based binaural noise reduction algorithms using partial noise estimation
  595. Interference statistics in a random mmWave ad hoc network
  596. Intonational phrase break prediction for text-to-speech synthesis using dependency relations
  597. Inverse Reinforcement Learning using Expectation Maximization in mixture models
  598. Investigating bias in non-parametric mutual information estimation
  599. Investigating online low-footprint speaker adaptation using generalized linear regression and click-through data
  600. Investigation of a parametric gain approach to single-channel speech enhancement
  601. Investigation of ensemble models for sequence learning
  602. Investigation of mixture splitting concept for training linear bottlenecks of deep neural network acoustic models
  603. Investigations on sequence training of neural networks
  604. Iterative randomized robust linear regression
  605. JFA modeling with left-to-right structure and a new backend for text-dependent speaker recognition
  606. Jammer forensics: Localization in peer to peer networks based on Q-learning
  607. Joint acoustic and spectral modeling for speech dereverberation using non-negative representations
  608. Joint covariance estimation with mutual linear structure
  609. Joint denoising and contrast enhancement of images using graph laplacian operator
  610. Joint design of multi-tap filters and power control for FBMC/OQAM based two-way decode-and-forward relaying systems in highly frequency selective channels
  611. Joint direction-of-arrival and frequency estimation without source enumeration
  612. Joint directional-positional multiplexing for light field acquisition by Kronecker compressed sensing
  613. Joint estimation of vocal tract and nasal tract area functions from speech waveforms via auto-regression moving-average modeling and a pole assignment method
  614. Joint group power allocation and prebeamforming for joint spatial-division multiplexing in multiuser massive MIMO systems
  615. Joint hot and cold clutter mitigation in the transmit beamspace-based MIMO radar
  616. Joint optimization of anatomical and gestural parameters in a physical vocal tract model
  617. Joint optimization of loudspeaker placement and radiation patterns for Sound Field Reproduction
  618. Joint time reversal and compressive sensing based localization algorithms for multiple-input multiple-output radars
  619. Joint time- and frequency-domain reshaping of room impulse responses
  620. Joint training of front-end and back-end deep neural networks for robust speech recognition
  621. K-medians clustering based ℓ1-PCA model
  622. KL-HMM based speaker diarization system for meetings
  623. Kernel Additive Modeling for interference reduction in multi-channel music recordings
  624. Kernel task-driven dictionary learning for hyperspectral image classification
  625. Kernel-based embeddings for large graphs with centrality constraints
  626. Knowledge Graph Inference for spoken dialog systems
  627. LOST-find: A spectral-space-time direct blind geolocalization algorithm
  628. Labelwalking nonnegative matrix factorization
  629. Language independent query-by-example spoken term detection using N-best phone sequences and partial matching
  630. Language model adaptation for academic lectures using character recognition result of presentation slides
  631. Language-independent voice passphrase verification
  632. Language-resource independent speech segmentation using cues from a spectrogram image
  633. Laplacian matrix learning for smooth graph signal representation
  634. Large dimensional analysis of Maronna's M-estimator with outliers
  635. Large region acoustic source mapping using movable arrays
  636. Large-scale sensor network localization via rigid subnetwork registration
  637. Large-scale speaker search using PLDA on mismatched conditions
  638. Large-scaleword representation features for improved spoken language understanding
  639. Lasso-based reverberation suppression in automatic speech Recognition
  640. Latent time-frequency component analysis: A novel pitch-based approach for singing voice separation
  641. Lattice FIR digital filter architectures using stochastic computing
  642. Learning acoustic frame labeling for speech recognition with recurrent neural networks
  643. Learning by weakly-connected adaptive agents
  644. Learning discriminative visual dictionary for natural scene categorization
  645. Learning feature mapping using deep neural network bottleneck features for distant large vocabulary speech recognition
  646. Learning interpretable classification rules using sequential rowsampling
  647. Learning joint features for color and depth images with Convolutional Neural Networks for object classification
  648. Learning mixed divergences in coupled matrix and tensor factorization models
  649. Learning shared rankings from mixtures of noisy pairwise comparisons
  650. Learning the sparsity basis in low-rank plus sparse model for dynamic MRI reconstruction
  651. Leveraging automatic speech recognition in cochlear implants for improved speech intelligibility under reverberation
  652. Leveraging valence and activation information via multi-task learning for categorical emotion recognition
  653. Librispeech: An ASR corpus based on public domain audio books
  654. Linear prediction based comfort noise generation in the EVS codec
  655. Linear support vector machines with normalizations
  656. Local All-Pass filters for optical flow estimation
  657. Local and global optimality of LP minimization for sparse recovery
  658. Local binary pattern orientation based face recognition
  659. Local metric learning for EEG-based personal identification
  660. Localization of a moving non-cooperative RF target in NLOS environment using RSS and AOA measurements
  661. Localized error detection for targeted clarification in a virtual assistant
  662. Location robust estimation of predictive Weibull parameters in short-term wind speed forecasting
  663. Location-aware object detection via coherent region grouping
  664. Logistic similarity metric learning for face verification
  665. Long short term memory neural network for keyboard gesture decoding
  666. Long short-term memory language models with additive morphological features for automatic speech recognition
  667. Lossless plenoptic image compression using adaptive block differential prediction
  668. Low bit rate high-quality MDCT audio coding of the 3GPP EVS standard
  669. Low delay LPC and MDCT-based audio coding in the EVS codec
  670. Low rank tensor deconvolution
  671. Low-complexity and robust coding mode decision in the EVS coder
  672. Low-complexity compressive sensing detection for multi-user spatial modulation systems
  673. Low-complexity minimum-SER channel equalization for OFDM underwater acoustic communications
  674. Low-complexity robust DOA estimation
  675. Low-complexity robust MISO downlink precoder optimization for the limited feedback case
  676. Low-latency list decoding of polar codes with double thresholding
  677. Low-latency sound-source-separation using non-negative matrix factorisation with coupled analysis and synthesis dictionaries
  678. Low-rank approximation-based distributed node-specific signal estimation in a fully-connected wireless sensor network
  679. Low-resource keyword search strategies for tamil
  680. Lowpass/bandpass signal reconstruction and digital filtering from nonuniform samples
  681. Lp-norm non-negative matrix factorization and its application to singing voice enhancement
  682. MASK+: Data-driven regions selection for acoustic fingerprinting
  683. MDCT audio coding with pulse vector quantizers
  684. MIST: L0 sparse linear regression with momentum
  685. ML estimation of population size when observing multiple fill levels in slotted Aloha
  686. Malware classification with recurrent networks
  687. Marginal Weiss-Weinstein bounds for discrete-time filtering
  688. Matching Musical Themes based on noisy OCR and OMR input
  689. Max-product dynamical systems and applications to audio-visual salient event detection in videos
  690. Maximal multiplicative spatial-spectral concentration on the sphere: Optimal basis
  691. Maximum a posteriori estimation of room impulse responses
  692. Maximum entropy property of discrete-time stable spline kernel
  693. Maximum expected achievable rate combining for limited feedback block-diagonalization
  694. Maximum likelihood approach to "informed" Sound Source Localization for Hearing Aid applications
  695. Maximum likelihood nonlinear transformations based on deep neural networks
  696. Mean square analysis of the CLMS and ACLMS for non-circular signals: The approximate uncorrelating transform approach
  697. Measure-transformed quasi maximum likelihood estimation with application to source localization
  698. Memory-aware i-vector extraction by means of sub-space factorization
  699. Meta-level tracking for gestural intent recognition
  700. Methods for applying dynamic sinusoidal models to statistical parametric speech synthesis
  701. Metric-Constrained Kernel Union of Subspaces
  702. Metrics in the space of high order proximity networks
  703. Metrics of grassmannian representation in reproducing kernel hilbert space for variational pattern analysis
  704. Micbots: Collecting large realistic datasets for speech and audio research using mobile robots
  705. Microphone array for increasing mutual information between sound sources and observation signals
  706. Microphone array position calibration in the frequency domain using a single unknown source
  707. Minimum Bayes risk signal detection for speech enhancement based on a narrowband DOA model
  708. Minimum information dominating set for critical sampling over graphs
  709. Mismatched filter design for radar waveforms by semidefinite relaxation
  710. Missing intensity restoration via adaptive selection of perceptually optimized subspaces
  711. Mixer-based subarray beamforming for sub-Nyquist sampling ultrasound architectures
  712. Mobile adaptive networks for pursuing multiple targets
  713. Mode Dependent Vector Quantization with a rate-distortion optimized codebook for residue coding in video compression
  714. Model-based parameters estimation of non-stationary signals using time warping and a measure of spectral concentration
  715. Model-distributed solution of regularized least-squares problem over sensor networks
  716. Model-order selection for analyzing correlation between two data sets using CCA with PCA preprocessing
  717. Modeling inter-node acoustic dependencies with Restricted Boltzmann Machine for distributed microphone array based BSS
  718. Modeling long temporal contexts in convolutional neural network-based phone recognition
  719. Modeling mutual influence of multimodal behavior in affective dyadic interactions
  720. Modelling acoustic feature dependencies with artificial neural networks: Trajectory-RNADE
  721. Modelling of complex signals using gaussian processes
  722. Modelling the decay of piano sounds
  723. Modified distributed iterative hard thresholding
  724. Modulation Wiener filter for improving speech intelligibility
  725. Modulation classification in MIMO fading channels via expectation maximization with non-data-aided initialization
  726. Modulation spectrum-constrained trajectory training algorithm for GMM-based Voice Conversion
  727. Monotone optimal policies in portfolio liquidation problems
  728. Motion compensation with higher order motion models for HEVC
  729. Moving sound source parameter estimation using a single microphone and signal extrema samples
  730. Multi-Modulus algorithms using hyperbolic and givens rotations for blind deconvolution of mimo systems
  731. Multi-basis adaptive neural network for rapid adaptation in speech recognition
  732. Multi-channel PSD estimators for speech dereverberation - A theoretical and experimental comparison
  733. Multi-channel linear prediction-based speech dereverberation with low-rank power spectrogram approximation
  734. Multi-channel speaker localization and separation using a model-based GSC and an inertial measurement unit
  735. Multi-frame factorisation for long-span acoustic modelling
  736. Multi-graph learning of spectral graph dictionaries
  737. Multi-instrument detection in polyphonic music using Gaussian Mixture based factorial HMM
  738. Multi-lingual speech recognition with low-rank multi-task deep neural networks
  739. Multi-parameter estimation for cognitive radar in compound Gaussian clutter
  740. Multi-resolution signal decomposition with time-domain spectrogram factorization
  741. Multi-scale Bayesian reconstruction of compressive X-ray image
  742. Multi-scale multi-lag channel estimation via linearization of training signal spectrum and sparse approximation
  743. Multi-sensor classification via sparsity-based representation with low-rank interference
  744. Multi-shift principal component analysis based primary component extraction for spatial audio reproduction
  745. Multi-source direction-of-arrival estimation in a reverberant environment using single acoustic vector sensor
  746. Multi-speaker modeling and speaker adaptation for DNN-based TTS synthesis
  747. Multi-task deep neural network acoustic models with model adaptation using discriminative speaker identity for whisper recognition
  748. Multi-task rank learning for image quality assessment
  749. Multi-view implicit transfer for person re-identification
  750. Multi-view indoor scene reconstruction from compressed through-wall radar measurements using a joint bayesian sparse representation
  751. Multicast beamforming with antenna selection using exact penalty approach
  752. Multichannel Wiener filtering via multichannel decorrelation
  753. Multichannel speech enhancement using MEMS microphones
  754. Multichannel transient acoustic signal classification using task-driven dictionary with joint sparsity and beamforming
  755. Multidimensional Ramanujan-sum expansions on nonseparable lattices
  756. Multimodal addressee detection in multiparty dialogue systems
  757. Multimodal arousal rating using unsupervised fusion technique
  758. Multipath exploitation in sparse scene recovery using sensing-through-wall distributed radar sensor configurations
  759. Multipitch estimation using a PLCA-based model: Impact of partial user annotation
  760. Multiple Early Termination for fast HEVC coding of UHD content
  761. Multiple constant multiplication implementations in near-threshold computing systems
  762. Multiple instance learning for breast MRI based on generic spatio-temporal features
  763. Multiple particle filtering with improved efficiency and performance
  764. Multiple source localization with moving co-prime arrays
  765. Multiple target track-before-detect in compound Gaussian clutter
  766. Multitask diffusion LMS with sparsity-based regularization
  767. Multiuser charging control in wireless power transfer via magnetic resonant coupling
  768. Multiuser cooperative transmission through superposition modulation based on braid coding
  769. Multivariate lattices for encrypted image processing
  770. Music separation guided by cover tracks: Designing the joint NMF model
  771. NMF-based blind source separation using a linear predictive coding error clustering criterion
  772. Narrow-range frequency estimation based on comprehensive optimization of DFT and interpolation
  773. Near-field sound propagation based on a circular and linear array combination
  774. Nearest neighbor based i-vector normalization for robust speaker recognition under unseen channel conditions
  775. Nearest neighbor discriminant analysis for language recognition
  776. Neighborhood regression for edge-preserving image super-resolution
  777. Nested generalized sidelobe canceller for joint dereverberation and noise reduction
  778. Network formation games based on conditional independence graphs
  779. Network infection source identification under the SIRI model
  780. Neural network joint modeling via context-dependent projection
  781. Neuron sparseness versus connection sparseness in deep neural network for large vocabulary speech recognition
  782. New post-processing techniques for low bit rate celp codecs
  783. Noise PSD estimation by logarithmic baseline tracing
  784. Noise cleaning and Gaussian modeling of smart phone photoplethysmogram to improve blood pressure estimation
  785. Noise reduced high dynamic range tone mapping using information content weights
  786. Noise reduction for screen content coding based on local histogram
  787. Noise robust estimation of the voice source using a deep neural network
  788. Noise robust integration for blind and non-blind reverberation time estimation
  789. Noise-shaping for closed-loop Multi-Channel Linear Prediction
  790. Noisy channel detection using the common annihilator with an application to electrocardiograms
  791. Non-linear acoustic echo cancellation using empirical mode decomposition
  792. Non-linear distortion reduction for a loudspeaker based on recursive source equalization
  793. Non-negative matrix factorisation incorporating greedy hellinger sparse coding applied to polyphonic music transcription
  794. Noncoherent sequence detection of orthogonally modulated signals in flat fading with log-linear complexity
  795. Nonconvex relaxation for Poisson intensity reconstruction
  796. Nonlinear regression using smooth Bayesian estimation
  797. Nonlinear, reduced order, distributed state estimation in microgrids
  798. Nonlocal means image denoising based on bidirectional principal component analysis
  799. Nonnegative matrix factorization with gradient vertex pursuit
  800. Nonuniformly sampled trivariate empirical mode decomposition
  801. Normalization of total variability matrix for i-vector/PLDA speaker verification
  802. Novel GCC-PHAT model in diffuse sound field for microphone array pairwise distance based calibration
  803. Novel audio features for capturing tempo salience in music recordings
  804. Novel autoregressive model based on adaptive window-extension and patch-geodesic distance for image interpolation
  805. Novel image classification based on integration of EEG and visual features via MSLPCCA
  806. Novel sound mixing method for voice and background music
  807. OOV Proper Name retrieval using topic and lexical context models
  808. Objective quality prediction for haptic texture signal compression
  809. Objective speech intelligibility assessment through comparison of phoneme class conditional probability sequences
  810. Objects co-segmentation: Propagated from simpler images
  811. Ocean acoustic waveguide invariant parameter estimation using tonal noise sources
  812. Ocrapose: An indoor positioning system using smartphone/tablet cameras and OCR-aided stereo feature matching
  813. On application of non-negative matrix factorization for ad hoc microphone array calibration from incomplete noisy distances
  814. On automatic drum transcription using non-negative matrix deconvolution and itakura saito divergence
  815. On clock synchronization for multi-microphone speech processing in wireless acoustic sensor networks
  816. On directivity factor of the first-order steerable differential microphone array
  817. On finding a subset of non-defective items from a large population using group tests: Recovery algorithms and bounds
  818. On frequency domain models for TDOA estimation
  819. On optimal mobile RSSI-sensor positioning for multi target tracking
  820. On optimal routing and power allocation for D2D communications
  821. On quantifying facial expression-related atypicality of children with Autism Spectrum Disorder
  822. On speech quality estimation of phase-aware single-channel speech enhancement
  823. On the Cramér-Rao lower bound under model mismatch
  824. On the broadband effect of remote stations in DPD algorithm
  825. On the complexity of information planning in Gaussian models
  826. On the design of the measurement matrix for Compressed Sensing based DOA estimation
  827. On the detection of abandoned objects with a moving camera using robust subspace recovery and sparse representation
  828. On the distributed acoustic sensing based on local time-frequency coherence analysis
  829. On the fly estimation of the sparsity degree in Compressed Sensing using sparse sensing matrices
  830. On the importance of modeling and robustness for deep neural network feature
  831. On the importance of using high resolution images, third level features and sequence of images for fingerprint spoof detection
  832. On the influence of microphone array geometry on HRTF-based Sound Source Localization
  833. On the potential for artificial bandwidth extension of bone and tissue conducted speech: a mutual information study
  834. On the preprocessing and postprocessing of HRTF individualization based on sparse representation of anthropometric features
  835. On the spectral growth of the polar representation of communication signals
  836. On the use of the tempogram to describe audio content and its application to Music structural segmentation
  837. On the von mises approximation for the distribution of the phase angle between two independent complex Gaussian vectors
  838. On transmit beamforming in MIMO radar with matrix completion
  839. On using heterogeneous data for vehicle-based speech recognition: A DNN-based approach
  840. One-formant vocal tract modeling for glottal pulse shape estimation
  841. Online adaptative zero-shot learning spoken language understanding using word-embedding
  842. Online computation of sparse representations of time varying stimuli using a biologically motivated neural network
  843. Online learning based on iterative projections in sum space of linear and Gaussian reproducing kernel Hilbert spaces
  844. Online local Gaussian process for tensor-variate regression: Application to fast reconstruction of limb movements from brain signal
  845. Online time-dependent clustering using probabilistic topic models
  846. Optically visualized sound field reconstruction based on sparse selection of point sound sources
  847. Optimal base station densities for cost-efficient multi-tier heterogeneous cellular networks
  848. Optimal beamforming on synthetic interference and noise for active processing of multiplet line arrays
  849. Optimal design and power allocation for multicarrier decode and forward relays
  850. Optimal design of directivity patterns for endfire linear microphone arrays
  851. Optimal error feedback filters for uniform quantizers at remote sensors
  852. Optimal geometry analysis for elliptic target localization by multistatic radar with independent bistatic channels
  853. Optimal graph laplacian regularization for natural image denoising
  854. Optimal sensor deployment for 3D AOA target localization
  855. Optimal single-channel noise reduction filtering matrices from the pearson correlation coefficient perspective
  856. Optimal spatial filtering for auditory steady-state response detection using high-density EEG
  857. Optimization for randomly described arrays based on geometry descriptors
  858. Optimization methods for sequence design with low autocorrelation sidelobes
  859. Optimization of plug-in electric vehicle charging with forecasted price
  860. Optimum decision fusion in cognitive wireless sensor networks with unknown users location
  861. Optimum discrete distributed beamforming for single group multicasting relay networks with relay selection
  862. Optimum node selection for protection under power grid state estimation
  863. Optimum phase-only discrete broadcast beamforming with antenna and user selection in interference limited cognitive radio networks
  864. Order-free spoken term detection
  865. Ordinal pyramid pooling for rotation invariant object recognition
  866. Outlier identification via randomized adaptive compressive sampling
  867. Overview of the EVS codec architecture
  868. PLDA-based diarization of telephone conversations
  869. Packet-loss concealment technology advances in EVS
  870. Parallel algorithms for large scale constrained tensor decomposition
  871. Parallel software implementation of recursive multidimensional digital filters for point-target detection in cluttered infrared scenes
  872. Parallelizable PARAFAC decomposition of 3-way tensors
  873. Parameter estimation for multiple scattering process on the sphere
  874. Parameter extraction for bass guitar sound models including playing styles
  875. Parameter generation algorithm considering Modulation Spectrum for HMM-based speech synthesis
  876. Parametric binaural rendering utilizing compact microphone arrays
  877. Parametric frugal sensing of Moving Average power spectra
  878. Paraphrastic recurrent neural network language models
  879. Particle Gibbs with refreshed backward simulation
  880. Particle filtering of ARMA processes of unknown order and parameters
  881. Particle filtering with observations in a manifold
  882. Patch-disagreement as away to improve K-SVD denoising
  883. Pattern based anomalous user detection in cognitive radio networks
  884. Pattern classification formulated as a missing data task: The audio genre classification case
  885. Pattern discovery from audio recordings by Variable Markov Oracle: A music information dynamics approach
  886. Pedestrian detection via PCA filters based convolutional channel features
  887. Pedestrian localization in moving platforms using dead reckoning, particle filtering and map matching
  888. Perceptual effect of reverberation on multi-microphone noise reduction for cochlear implants
  889. Performance analysis of music in the presence of modeling errors due to the spatial distributions of sources
  890. Performance analysis of spatial smoothing schemes in the context of large arrays
  891. Performance analysis of the covariance subtraction method for relative transfer function estimation and comparison to the covariance whitening method
  892. Performance estimation for tensor CP decomposition with structured factors
  893. Periodic RF transmitter geolocation using a mobile receiver
  894. Periodic non-uniform sampling for FRI signals
  895. Persistent topology of decision boundaries
  896. Phase recovery for time of arrival estimation in the presence of interference
  897. Phase recovery from a Bayesian point of view: The variational approach
  898. Phase recovery in NMF for audio source separation: An insightful benchmark
  899. Phase transition of joint-sparse recovery from multiple measurements via convex optimization
  900. Phase transitions in spectral community detection of large noisy networks
  901. Phase-based detection of intentional state for asynchronous brain-computer interface
  902. Phase-optimized K-SVD for signal extraction from underdetermined multichannel sparse mixtures
  903. Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks
  904. Phonological vocoding using artificial neural networks
  905. Photo-real talking head with deep bidirectional LSTM
  906. Piano music transcription modeling note temporal evolution
  907. Pitch and TDOA-based localization of acoustic sources with distributed arrays
  908. Pitch estimation and tracking with harmonic emphasis on the acoustic spectrum
  909. Plane sweep method for optimal line fitting in track-before-detect
  910. Planted clique detection below the noise floor using low-rank sparse PCA
  911. Polynomial-phase signal direction-finding and source-tracking with a single acoustic vector sensor
  912. Posture-invariant ECG recognition with posture detection
  913. Potts model parameter estimation in Bayesian segmentation of piecewise constant images
  914. Precise error analysis of the LASSO
  915. Precoder and equalizer design for multi-user MIMO FBMC/OQAM with highly frequency selective channels
  916. Predicting next speaker based on head movement in multi-party meetings
  917. Privacy constrained energy management for self-interested microgrids
  918. Privacy-preserving Query-by-Example Speech Search
  919. Probabilistic features for connecting eye gaze to spoken language understanding
  920. Probability density function estimation by positive quartic C2-spline functions
  921. Prof-Life-Log: Analysis and classification of activities in daily audio streams
  922. Prosody generation using frame-based Gaussian process regression and classification for statistical parametric speech synthesis
  923. Proximal diffusion for stochastic costs with non-differentiable regularizers
  924. Pseudo-coherence-based MVDR beamformer for speech enhancement with ad hoc microphone arrays
  925. QUESST2014: Evaluating Query-by-Example Speech Search in a zero-resource setting with real-life queries
  926. Quality estimation for asr k-best list rescoring in spoken language translation
  927. Quantifying EDA synchrony through joint sparse representation: A case-study of couples' interactions
  928. Quantile analysis of image sensor noise distribution
  929. Quantized fuzzy LBP for face recognition
  930. Quantized matrix completion for low rank matrices
  931. Quasi-rectilinear (MSK, GMSK, OQAM) co-channel interference mitigation by three inputs widely linear fresh filtering
  932. Query-by-example keyword spotting using long short-term memory networks
  933. Quickest detection of short-term voltage instability with PMU measurements
  934. Raking echoes in the time domain
  935. Ramanujan filter banks for estimation and tracking of periodicities
  936. Random matrix theory inspired passive bistatic radar detection with noisy reference signal
  937. Random projection and multiscale wavelet leader based anomaly detection and address identification in internet traffic
  938. Random sequential scheduling for wireless D2D communications
  939. Ranging without time stamps exchanging
  940. Rapid: Rapidly accelerated proximal gradient algorithms for convex minimization
  941. Rate control for lossless region of interest coding in HEVC intra-coding with applications to digital pathology images
  942. Rational consumer behavior models in smart pricing
  943. Real-time independent vector analysis with Student's t source prior for convolutive speech mixtures
  944. Real-time multiple DOA estimation of speech sources in wireless acoustic sensor networks
  945. Real-time robust formant tracking system using a phase equalization-based autoregressive exogenous model
  946. Real-time self-tracking in the Internet of Things
  947. Recovering signals from the Short-Time Fourier Transform magnitude
  948. Recurrent neural network language model training with noise contrastive estimation for speech recognition
  949. Recurrent neural network language model with structured word embeddings for speech recognition
  950. Reduced vowel space is a robust indicator of psychological distress: A cross-corpus analysis
  951. Reduced-rank condensed filter dictionaries for inter-picture prediction
  952. Reduced-rank modeling of time-varying spectral patterns for supervised source separation
  953. Reducing communication overhead in distributed learning by an order of magnitude (almost)
  954. Reducing quantization error in low-energy FIR filter accelerators
  955. Redundancy analysis of behavioral coding for couples therapy and improved estimation of behavior from noisy annotations
  956. Reference-distance estimation approach for TDOA-based source and sensor localization
  957. Region-based depth map coding using a 3D scene representation
  958. Regret bounds of a distributed saddle point algorithm
  959. Regularization of context-dependent deep neural networks with context-independent multi-task training
  960. Regularized canonical correlations for sensor data clustering
  961. Regularizing DNN acoustic models with Gaussian stochastic neurons
  962. Relative group sparsity for non-negative matrix factorization with application to on-the-fly audio source separation
  963. Removing data with noisy responses in regression analysis
  964. Representation and modeling of spherical harmonics manifold for source localization
  965. Representation models in single channel source separation
  966. Residual noise control using a parametric multichannel Wiener filter
  967. Restricted Boltzmann Machine supervectors for speaker recognition
  968. Risk-averse online learning under mean-variance measures
  969. Robot audition: Its rise and perspectives
  970. Robust DOA estimation of heavily noisy gunshot signals
  971. Robust and computationally efficient diffusion-based classification in distributed networks
  972. Robust and reliable audio watermarking based on phase coding
  973. Robust audio surveillance using spectrogram image texture feature
  974. Robust beamformer and artificial noises for MISO wiretap channels with multiple eavesdroppers
  975. Robust binary hypothesis testing under contaminated likelihoods
  976. Robust electroencephalogram channel set for person authentication
  977. Robust estimation of structured covariance matrix for heavy-tailed distributions
  978. Robust excitation-based features for Automatic Speech Recognition
  979. Robust joint beamforming and artificial noise design for amplify-and-forward multi-antenna relay systems
  980. Robust linear spectral unmixing using outlier detection
  981. Robust localisation of multiple speakers exploiting head movements and multi-conditional training of binaural cues
  982. Robust long term neural signal decoding by estimating unobserved features
  983. Robust microphone placement for source localization from noisy distance measurements
  984. Robust minimum variance beamforming under distributional uncertainty
  985. Robust overlapped speech detection and its application in word-count estimation for Prof-Life-Log data
  986. Robust precoding design for multibeam downlink satellite channel with phase uncertainty
  987. Robust sound event recognition using convolutional neural networks
  988. Robust speech processing using ARMA spectrogram models
  989. Robust statistical process control in Block-RDT framework
  990. Robust transmit beampattern design for uniform linear arrays using correlated LFM waveforms
  991. Robust unsupervised detection of human screams in noisy acoustic environments
  992. Robust widely linear beamformer based on a projection constraint
  993. SAS: A speaker verification spoofing database containing diverse attacks
  994. SINR loss of the dominant mode rejection beamformer
  995. SKA correlators and beamformers
  996. SNR maximization hashing for learning compact binary codes
  997. SOBM - a binary mask for noisy speech that optimises an objective intelligibility metric
  998. Salient object detection via background contrast
  999. Sampling smooth spatio-temporal physical fields: When will the aliasing error increase with time?
  1000. Sampling spherical finite rate of innovation signals

Looking for submission deadlines instead? See the conference deadline calendar.