← All conferences

ICASSP 2018 Accepted Papers

The full list of 1,393 papers accepted at ICASSP 2018 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

  1. (Almost) Zero-Shot Cross-Lingual Spoken Language Understanding
  2. 2.5D Multizone Reproduction Using Weighted Mode Matching
  3. 2D Vector Map Reversible Data Hiding with Topological Relation Preservation
  4. 3-D CNN Models for Far-Field Multi-Channel Speech Recognition
  5. 3D Exterior Soundfield Reproduction Using a Planar Loudspeaker Array
  6. 3D Image Reconstruction from Multi-Focus Microscope: Axial Super-Resolution and Multiple-Frame Processing
  7. 3D Mouth Tracking from a Compact Microphone Array Co-Located with a camera
  8. 3D-Hog Embedding Frameworks for Single and Multi-Viewpoints Action Recognition Based on Human Silhouettes
  9. A 203 FPS VLSI Architecture of Improved Dense Trajectories for Real-Time Human Action Recognition
  10. A 320M Pixel/S Vlsi Architecture Design of Weighted Mode Filter for 4K Ultra-Hd Depth Upsampling
  11. A Bayesian Framework to Optimize Double Band Spectra Spatial Filters for Motor Imagery Classification
  12. A Bayesian Hierarchical Model for Speech Enhancement
  13. A Bayesian Model for Joint Unmixing and Robust Classification of Hyperspectral Images
  14. A Capitalist Scheme for Energy Management in Inferential Sensor Networks
  15. A Captcha Design Based on Visual Reasoning
  16. A Casa Approach to Deep Learning Based Speaker-Independent Co-Channel Speech Separation
  17. A Cascaded Framework for Model-Based 3D Face Reconstruction
  18. A Comparison of Recent Waveform Generation and Acoustic Modeling Methods for Neural-Network-Based Speech Synthesis
  19. A Complete End-to-End Speaker Verification System Using Deep Neural Networks: From Raw Signals to Verification Result
  20. A Compressive Sensing-Based Active User and Symbol Detection Technique for Massive Machine-Type Communications
  21. A Constant Step Stochastic Douglas-Rachford Algorithm with Application to non Separable Regularizations
  22. A Conversational Neural Language Model for Speech Recognition in Digital Assistants
  23. A Corrective Learning Approach for Text-Independent Speaker Verification
  24. A Coupled Compressive Sensing Scheme for Unsourced Multiple Access
  25. A Deep Dictionary Model for Image Super-Resolution
  26. A Deep Encoder-Decoder Networks for Joint Deblurring and Super-Resolution
  27. A Deep Learning Based Alternative to Beamforming Ultrasound Images
  28. A Deep Learning Based No-Reference Image Quality Assessment Model for Single-Image Super-Resolution
  29. A Deep Neural Network Approach for Time-Of- Arrival Estimation in Multipath Channels
  30. A Deep Neural Network Based Method of Source Localization in a Shallow Water Environment
  31. A Deep Reinforcement Learning Framework for Identifying Funny Scenes in Movies
  32. A Deeper Look at Gaussian Mixture Model Based Anti-Spoofing Systems
  33. A Deeply-Recursive Convolutional Network For Crowd Counting
  34. A Dimension-Independent Discriminant Between Distributions
  35. A Discrete Signal Processing Framework for Set Functions
  36. A Discriminatively Learned Feature Embedding Based on Multi-Loss Fusion For Person Search
  37. A Distance-Based Formulation for Sampling Signals on Graphs
  38. A Diverse Large-Scale Dataset for Evaluating Rebroadcast Attacks
  39. A Dynamic Latent Variable Model for Source Separation
  40. A Family of Matrices for Generating Hermite-Gaussian-Like DFT Eigenvectors
  41. A Fast and Memory-Efficient Algorithm for Robust PCA (MEROP)
  42. A Feature Fusion Method Based on Extreme Learning Machine for Speech Emotion Recognition
  43. A Flexible Dirty Model Dictionary Learning Approach for Classification
  44. A Fully Convolutional Tri-Branch Network (FCTN) for Domain Adaptation
  45. A Generalized Uncorrelated Ridge Regression with Nonnegative Labels for Unsupervised Feature Selection
  46. A Generative Adversarial Network Based Framework for Unsupervised Visual Surface Inspection
  47. A Generative Auditory Model Embedded Neural Network for Speech Processing
  48. A Graph-CNN for 3D Point Cloud Classification
  49. A Greedy Pursuit Algorithm for Separating Signals from Nonlinear Compressive Observations
  50. A Hybrid Approach to Combining Conventional and Deep Learning Techniques for Single-Channel Speech Enhancement and Recognition
  51. A Hybrid Dictionary Approach for Distributed Kernel Adaptive Filtering in Diffusion Networks
  52. A Hybrid Neural Network Based on the Duplex Model of Pitch Perception for Singing Melody Extraction
  53. A Joint Detection and Reconstruction Method for Blind Graph Signal Recovery
  54. A Joint Multi-Task Learning Framework for Spoken Language Understanding
  55. A Joint Perspective of Periodically Excited Efficient NLMS Algorithm and Inverse Cyclic Convolution
  56. A Joint Separation-Classification Model for Sound Event Detection of Weakly Labelled Data
  57. A Joint Source Channel Arithmetic Map Decoder Using Probabilistic Relations Among Intra Modes in Predictive Video Compression
  58. A Joint Target Localization and Classification Framework for Sensor Networks
  59. A Large-Scale Study of Language Models for Chord Prediction
  60. A Learning Algorithm with Compression-Based Regularization
  61. A Light-Weight Multimodal Framework for Improved Environmental Audio Tagging
  62. A Low Power Hardware Implementation of Multi-Object DPM Detector for Autonomous Driving
  63. A Low-Complexity Video Encoder for Equirectangular Projected 360 Video Content
  64. A Matrix Completion Approach for Wall-Clutter Mitigation in Compressive Radar Imaging of Indoor Targets
  65. A Modified Signal Phase Unwrapping Algorithm for Range Estimation
  66. A Motion Aided Merge Mode For Hevc
  67. A Multi-Camera Deep Neural Network for Detecting Elevated Alertness in Drivers
  68. A Multi-Resolution Approach to Complexity Reduction in Tomographic Reconstruction
  69. A Multi-Seed 3D Local Graph Matching Model for Tracking of Densely Packed Cells
  70. A Natural Shape-Preserving Stereoscopic Image Stitching
  71. A New Proximal Method for Joint Image Restoration and Edge Detection with the Mumford-Shah Model
  72. A Nonconvex Variational Approach for Robust Graphical Lasso
  73. A Nonlinear 3D Geometric Tongue Model
  74. A Novel Crowd-Resilient Visual Localization Algorithm Via Robust Pca Background Extraction
  75. A Novel Ego-Noise Suppression Algorithm for Acoustic Signal Enhancement in Autonomous Systems
  76. A Novel Image-Specific Transfer Approach for Prostate Segmentation in MR Images
  77. A Novel Joint Radar and Communication System Based on Randomized Partition of Antenna Array
  78. A Novel LSTM-Based Speech Preprocessor for Speaker Diarization in Realistic Mismatch Conditions
  79. A Novel Learnable Dictionary Encoding Layer for End-to-End Language Identification
  80. A Novel Method for Human Bias Correction of Continuous- Time Annotations
  81. A Novel Selective Active Noise Control Algorithm to Overcome Practical Implementation Issue
  82. A Novel Semantic Attribute-Based Feature for Image Caption Generation
  83. A Novel Thresholding Technique for the Denoising of Multicomponent Signals
  84. A Parallel Best-Response Algorithm with Exact Line Search for Nonconvex Sparsity-Regularized Rank Minimization
  85. A Parallel Fusion Approach to Piano Music Transcription Based on Convolutional Neural Network
  86. A Parametric Approach for Classification of Distortions in Pathological Voices
  87. A Penalized Method for the Predictive Limit of Learning
  88. A Practical Guide to Multi-Image Alignment
  89. A Pragmatic Authentication System Using Electroencephalography Signals
  90. A Priori SNR Estimation Using Discriminative Non-Negative Matrix Factorization
  91. A Pruned Rnnlm Lattice-Rescoring Algorithm for Automatic Speech Recognition
  92. A Quaternion Kernel Minimum Error Entropy Adaptive Filter
  93. A Random Matrix and Concentration Inequalities Framework for Neural Networks Analysis
  94. A Refined Analysis of the Gap Between Expected Rate for Partial Csit and the Massive Mimo Rate Limit
  95. A Reliable Video Storage Architecture in Hybrid SLC/MLC Nand Flash
  96. A Revisit of Action Detection Using Improved Trajectories
  97. A Robust Change Detector for Highly Heterogeneous Multivariate Images
  98. A Robust Event-Triggered Consensus Strategy for Linear Multi-Agent Systems with Uncertain Network Topology
  99. A Robust Hierarchical Qp Setting for Screen Content Coding
  100. A Robust Machine Learning Method for Cell-Load Approximation in Wireless Networks
  101. A Rotation-Invariant Convolutional Neural Network for Image Enhancement Forensics
  102. A Second-Order Variational Framework for Joint Depth Map Estimation and Image Dehazing
  103. A Shuffled-Based Iterative Demodulation and Decoding Scheme for Ldpc Coded Flash Memory
  104. A Simple Cepstral Domain DNN Approach to Artificial Speech Bandwidth Extension
  105. A Simple and Effective Framework for a Priori SNR Estimation
  106. A Single-Channel Noise Reduction Filtering/Smoothing Technique in the Time Domain
  107. A Sparse Coding Framework for Gaze Prediction in Egocentric Video
  108. A Statistical Signal Processing Approach to Clustering over Compressed Data
  109. A Stem Reu Site on the Integrated Design of Sensor Devices and Signal Processing Algorithms
  110. A Study of All-Convolutional Encoders for Connectionist Temporal Classification
  111. A Study of Noise PSD Estimators for Single Channel Speech Enhancement
  112. A Study of Training Targets for Deep Neural Network-Based Speech Enhancement Using Noise Prediction
  113. A Supervised Air-Tissue Boundary Segmentation Technique in Real-Time Magnetic Resonance Imaging Video Using a Novel Measure of Contrast and Dynamic Programming
  114. A Supervised Approach to Global Signal-to-Noise Ratio Estimation for Whispered and Pathological Voices
  115. A Supervised Stdp-Based Training Algorithm for Living Neural Networks
  116. A Tensor Decomposition Technique for Source Localization from Multimodal Data
  117. A Time-Restricted Self-Attention Layer for ASR
  118. A Time-Weighted Method for Predicting the Intelligibility of Speech in the Presence of Interfering Sounds
  119. A Toeplitz-Tyler Estimation of the Model Order in Large Dimensional Regime
  120. A Triplet-Loss Embedded Deep Regressor Network for Estimating Blood Pressure Changes Using Prosodic Features
  121. A Two-Layer Reinforcement Learning Solution for Energy Harvesting Data Dissemination Scenarios
  122. A Unified Approach to Generating Sound Zones Using Variable Span Linear Filters
  123. A Unified Estimator for Source Positioning and DOA Estimation Using AOA
  124. A Wavenet for Speech Denoising
  125. A Weighted Least Squares Beam Shaping Technique for Sound Field Control
  126. ADA-PT: An Adaptive Parameter Tuning Strategy Based on the Weighted Stein Unbiased Risk Estimator
  127. ASR Performance Prediction on Unseen Broadcast Programs Using Convolutional Neural Networks
  128. Accelerated Image Reconstruction for Nonlinear Diffractive Imaging
  129. Accelerating Recurrent Neural Network Language Model Based Online Speech Recognition System
  130. Accent Conversion Using Phonetic Posteriorgrams
  131. Accounting for Room Acoustics in Audio-Visual Multi-Speaker Tracking
  132. Achievable Rate Maximization by Passive Intelligent Mirrors
  133. Achieving Accompanying Beampattern Peak for High-Speed Users Via Frequency Diverse Array
  134. Acoustic Analysis and Assessment of the Knee in Osteoarthritis During Walking
  135. Acoustic Feature Learning Using Cross-Domain Articulatory Measurements
  136. Acoustic Modeling of Speech Waveform Based on Multi-Resolution, Neural Network Signal Processing
  137. Acoustic Reflector Localization and Classification
  138. Acoustic Scene Classification Using Discrete Random Hashing for Laplacian Kernel Machines
  139. Acoustic-to-Word Attention-Based Model Complemented with Character-Level CTC-Based Model
  140. Active Anomaly Detection in Heterogeneous Processes
  141. Active Camera Relocalization with RGBD Camera from a Single 2D Image
  142. Active Covariance Estimation by Random Sub-Sampling of Variables
  143. Active Occlusion Cancellation with Hear-Through Equalization for Headphones
  144. Adaptation of an Expressive Single Speaker Deep Neural Network Speech Synthesis System
  145. Adaptive Bayesian Channel Gain Cartography
  146. Adaptive Clustering Algorithm for Cooperative Spectrum Sensing in Mobile Environments
  147. Adaptive Coding of Non-Negative Factorization Parameters with Application to Informed Source Separation
  148. Adaptive Noise Canceller with Snr Estimate Switchover for Stepsize Control
  149. Adaptive Parameters Adjustment for Group Reweighted Zero-Attracting LMS
  150. Adaptive Permutation Invariant Training with Auxiliary Information for Monaural Multi-Talker Speech Recognition
  151. Adaptive STFT with Chirp-Modulated Gaussian Window
  152. Adaptive Sparse Array Reconfiguration based on Machine Learning Algorithms
  153. Adaptive Travel Time Tomography with Local Sparsity
  154. Adaptive Visual Target Tracking Based on Label Consistent K-Svd Sparse Coding and Kernel Particle Filter
  155. Advanced LSTM: A Study About Better Time Dependency Modeling in Emotion Recognition
  156. Advancing Acoustic-to-Word CTC Model
  157. Advancing Connectionist Temporal Classification with Attention Modeling
  158. Adversarial Advantage Actor-Critic Model for Task-Completion Dialogue Policy Learning
  159. Adversarial Learning of Raw Speech Features for Domain Invariant Speech Recognition
  160. Adversarial Multi-Agent Target Tracking with Inexact Online Gradient Descent
  161. Adversarial Multilingual Training for Low-Resource Speech Recognition
  162. Adversarial Semi-Supervised Audio Source Separation Applied to Singing Voice Extraction
  163. Adversarial Teacher-Student Learning for Unsupervised Domain Adaptation
  164. Affine-Projection Least-Mean-Magnitude-Phase Algorithms Using a Posteriori Updates
  165. Alpha-Stable Low-Rank Plus Residual Decomposition for Speech Enhancement
  166. Alternating Minimization Approach for Identification of Piecewise Continuous Hammerstein Systems
  167. Alternative Objective Functions for Deep Clustering
  168. Altitude Measurement of Low-Angle Target Under Complex Terrain Environment for Meter-Wave Radar
  169. An Adaptive Combination Rule for Diffusion LMS Based on Consensus Propagation
  170. An Algorithm for Multi Subject Fmri Analysis Based on the SVD and Penalized Rank-1 Matrix Approximation
  171. An Analysis of Incorporating an External Language Model into a Sequence-to-Sequence Model
  172. An Analytical Method to Determine Minimum Per-Layer Precision of Deep Neural Networks
  173. An Architecture for Self -Aware IOT Applications
  174. An Attenuation Adapted Pulse Compression Technique to Enhance the Bandwidth and the Resolution Using Ultrafast Ultrasound Imaging
  175. An Efficient Deep Convolutional Laplacian Pyramid Architecture for Cs Reconstruction At Low Sampling Ratios
  176. An Efficient Residual Echo Suppression for Multi-Channel Acoustic Echo Cancellation Based on the Frequency-Domain Adaptive Kalman Filter
  177. An Efficient Target Localization Estimator from Bistatic Range and Tdoa Measurements in Multistatic Radar
  178. An End-To-End Siamese Convolutional Neural Network for Loop Closure Detection in Visual Slam System
  179. An End-to-End Approach to Joint Social Signal Detection and Automatic Speech Recognition
  180. An End-to-End Language-Tracking Speech Recognizer for Mixed-Language Speech
  181. An Ensemble Framework of Voice-Based Emotion Recognition System for Films and TV Programs
  182. An Ensemble Learning Approach to Detect Epileptic Seizures from Long Intracranial EEG Recordings
  183. An Ensemble Learning Method Based on Random Subspace Sampling for Palmprint Identification
  184. An Event-Triggered Average Consensus Algorithm with Performance Guarantees for Distributed Sensor Networks
  185. An Experimental Analysis of the Power Consumption of Convolutional Neural Networks for Keyword Spotting
  186. An Improved Doa Estimator Based on Partial Relaxation Approach
  187. An Improved Initialization for Low-Rank Matrix Completion Based on Rank-L Updates
  188. An Improved Iterative Algorithm for Band-Limited Signal Extrapolation on the Sphere
  189. An Investigation of Noise Shaping with Perceptual Weighting for Wavenet-Based Speech Generation
  190. An Investigation of Subband Wavenet Vocoder Covering Entire Audible Frequency Range with Limited Acoustic Features
  191. An Investigation of a Knowledge Distillation Method for CTC Acoustic Models
  192. An Iterative Approach for Shadow Removal in Document Images
  193. An Open-Source Speaker Gender Detection Framework for Monitoring Gender Equality
  194. An Unsupervised Anomalous Event Detection Framework with Class Aware Source Separation
  195. An Upper-Bound on the Required Size of a Neural Network Classifier
  196. An 𝓁0 Solution to Sparse Approximation Problems with Continuous Dictionaries
  197. Analysis and Optimization of Aperture Design in Computational Imaging
  198. Analysis of Multilingual Blstm Acoustic Model on Low and High Resource Languages
  199. Anatomy-Guided Inverse-Gradient Susceptibility Artifact Correction Method for High-Resolution FMRI
  200. Angle Dependent Match Filter Design for Circulating Code
  201. Anisotropic Total Variation Regularized Low-Rank Tensor Completion Based On Tensor Nuclear Norm for Color Image Inpainting
  202. Anscombe Meets Hough: Noise Variance Stablization Via Parametric Model Estimation
  203. Antenna Selection for Large-Scale Mimo Systems with Low-Resolution Adcs
  204. Aphash: Anchor-Based Probability Hashing for Image Retrieval
  205. Application of Progressive Neural Networks for Multi-Stream Wfst Combination in One-Pass Decoding
  206. Applying Multitask Learning to Acoustic-Phonemic Model for Mispronunciation Detection and Diagnosis in L2 English Speech
  207. Approximate Belief Propagation Decoder for Polar Codes
  208. Articulatory Information and Multiview Features for Large Vocabulary Continuous Speech Recognition
  209. Assessing Cross-Dependencies Using Bivariate Multifractal Analysis
  210. Asymmetric Dct-Jnd for Luminance Adaptation Effects: an Application To Perceptual Video Coding in Mv-Hevc
  211. Asymptotic Signal Detection Rates with 1-Bit Array Measurements
  212. Asynchronous Blind Network Division Multiple Access
  213. Attention-Based Dialog State Tracking for Conversational Interview Coaching
  214. Attention-Based End-to-End Speech Recognition on Voice Search
  215. Attention-Based LSTM for Psychological Stress Detection from Spoken Language Using Distant Supervision
  216. Attention-Based Models for Text-Dependent Speaker Verification
  217. Attitude Classification in Adjacency Pairs of a Human-Agent Interaction with Hidden Conditional Random Fields
  218. Audio Set Classification with Attention Model: A Probabilistic Perspective
  219. Audio Source Separation with Magnitude Priors: The Beads Model
  220. Audio Style Transfer
  221. Audio-Visual Conversation Analysis by Smart Posterboard and Humanoid Robot
  222. Audio-Visual Person Recognition in Multimedia Data From the Iarpa Janus Program
  223. Augmented Data and Improved Noise Residual-Based CNN for Printer Source Identification
  224. Augmented Latent Dirichlet Allocation (Lda) Topic Model with Gaussian Mixture Topics
  225. Augmenting Classrooms with AI for Personalized Education
  226. Autoencoder Based Image Compression: Can the Learning be Quantization Independent?
  227. Autoencoder Inspired Unsupervised Feature Selection
  228. Automated Detection of High FDG Uptake Regions in CT Images
  229. Automatic Bird Vocalization Identification Based on Fusion of Spectral Pattern and Texture Features
  230. Automatic Conflict Detection in Police Body-Worn Audio
  231. Automatic Motion Artifact Detection for Whole-Body Magnetic Resonance Imaging
  232. Automatic Music Transcription Leveraging Generalized Cepstral Features and Deep Learning
  233. Automatic Segmentation and Cardiopathy Classification in Cardiac Mri Images Based on Deep Neural Networks
  234. Automatic Shrinkage Tuning Robust to Input Correlation for Sparsity-Aware Adaptive Filtering
  235. Automatic Speech Assessment for Aphasic Patients Based on Syllable-Level Embedding and Supra-Segmental Duration Features
  236. Automatic Temporal Segmentation of Hand Movements for Hand Positions Recognition in French Cued Speech
  237. Automatically Linking Digital Signal Processing Assessment Questions to Key Engineering Learning Outcomes
  238. B-Spline Pdf: A Generalization of Histograms to Continuous Density Models for Generative Audio Networks
  239. BSS Eval or Peass? Predicting the Perception of Singing-Voice Separation
  240. Bandlimited Spatiotemporal Field Sampling with Location and Time Unaware Mobile Sensors
  241. Bayesian Anisotropic Gaussian Model for Audio Source Separation
  242. Bayesian Generative Model Based on Color Histogram of Oriented Phase and Histogram of Oriented Optical Flow for Rare Event Detection in Crowded Scenes
  243. Bayesian Inference for Multi-Line Spectra in Linear Sensor Array
  244. Bayesian Models for Unit Discovery on a Very Low Resource Language
  245. Bayesian Sparse Signal Detection Exploiting Laplace Prior
  246. Beamforming Design for Full-Duplex Cellular and Mimo Radar Coexistence: A Rate Maximization Approach
  247. Being Low-Rank in the Time-Frequency Plane
  248. Benchmarking Uncertainty Estimates with Deep Reinforcement Learning for Dialogue Policy Optimisation
  249. Binaural Rendering of Dynamic Head and Sound Source Orientation Using High-Resolution HRTF and Retarded Time
  250. Binaural Spectral Complexity Reduction of Music Signals for Cochlear Implant Listeners
  251. Binaural Speech Source Localization Using Template Matching of Interaural Time Difference Patterns
  252. Bindctnet: A Simple Binary Dct Network for Image Classification
  253. Birdvox-Full-Night: A Dataset and Benchmark for Avian Flight Call Detection
  254. Bitwise Neural Networks for Efficient Single-Channel Source Separation
  255. Bitwise Source Separation on Hashed Spectra: An Efficient Posterior Estimation Scheme Using Partial Rank Order Metrics
  256. Blind Bandwidth Extension Based on Convolutional and Recurrent Deep Neural Networks
  257. Blind Calibration for Acoustic Vector Sensor Arrays
  258. Blind Estimation of the Speech Transmission Index for Speech Quality Prediction
  259. Blind Image Deblurring Via Reweighted Graph Total Variation
  260. Blind Image Quality Assessment Based on Visuo-Spatial Series Statistics
  261. Blind Source Separation Using Mixtures of Alpha-Stable Distributions
  262. Block-Coordinate Proximal Algorithms for Scale-Free Texture Segmentation
  263. Boosting Noise Robustness of Acoustic Model via Deep Adversarial Training
  264. Boundary Objectness Network for Object Detection and Localization
  265. Breast Density Classification with Deep Convolutional Neural Networks
  266. Bridgenets: Student-Teacher Transfer Learning Based on Recursive Neural Networks and Its Application to Distant Speech Recognition
  267. Building Competitive Direct Acoustics-to-Word Models for English Conversational Speech Recognition
  268. CBLDNN-Based Speaker-Independent Speech Separation Via Generative Adversarial Training
  269. COMPASS: Coding and Multidirectional Parameterization of Ambisonic Sound Scenes
  270. CPD Updating Using Low-Rank Weights
  271. CRoss-lingual and Multilingual Speech Emotion Recognition on English and French
  272. CTC Loss Function with a Unit-Level Ambiguity Penalty
  273. Calibrating Cameras in Poor-Conditioned Pitch-Based Sports Games
  274. Can you Find a Face in a HEVC Bitstream?
  275. Capturing Shared and Individual Information in fMRI Data
  276. Cascade: Channel-Aware Structured Cosparse Audio Declipper
  277. Catseyes: Categorizing Seismic Structures with Tessellated Scattering Wavelet Networks
  278. Cell Subclass Identification in Single-Cell RNA-Sequencing Data Using Orthogonal Nonnegative Matrix Factorization
  279. Change-Point Detection of Gaussian Graph Signals with Partial Information
  280. Channel Dependent Codebook Design in Spatial Modulation
  281. Channel Dependent Mutual Information in Index Modulations
  282. Characterizing Performance of Speaker Diarization Systems on Far-Field Speech Using Standard Methods
  283. Classification of Corals in Reflectance and Fluorescence Images Using Convolutional Neural Network Representations
  284. Classification vs. Regression in Supervised Learning for Single Channel Speaker Count Estimation
  285. Classifier Cascade to Aid in Detection of Epileptiform Transients in Interictal EEG
  286. Classifying Pump-Probe Images of Melanocytic Lesions Using the WEYL Transform
  287. Cloud Radio Access Network with Optimized Base-Station Caching
  288. Cluster-Based Point Cloud Coding with Normal Weighted Graph Fourier Transform
  289. Clustering of Data with Missing Entries
  290. Clustering-Guided Gp-Ucb for Bayesian Optimization
  291. Coarray Interpolation-Based Coprime Array Doa Estimation Via Covariance Matrix Reconstruction
  292. Cognitive Analysis of Working Memory Load from Eeg, by a Deep Recurrent Neural Network
  293. Coherence Bounds for Sensing Matrices in Spherical Harmonics Expansion
  294. Coherent Time Reversal Sub-Array Processing for Microwave Breast Imaging
  295. Color Affine Subspace Pursuit for Color Artifact Removal
  296. Combining Acoustic Embeddings and Decoding Features for End-of-Utterance Detection in Real-Time Far-Field Speech Recognition Systems
  297. Combining Multiple Deep Features for Glaucoma Classification
  298. Combining Range and Direction for Improved Localization
  299. Common and Individual Feature Extraction Using Tensor Decompositions: a Remedy for the Curse of Dimensionality?
  300. Community Detection from Low-Rank Excitations of a Graph Filter
  301. Comparative Evaluations of Various Factored Deep Convolutional Rnn Architectures for Noise Robust Speech Recognition
  302. Comparing the Influence of Depth and Width of Deep Neural Network Based on Fixed Number of Parameters for Audio Event Detection
  303. Comparison of Speech Tasks for Automatic Classification of Patients with Amyotrophic Lateral Sclerosis and Healthy Subjects
  304. Complementary Complex-Valued Spectrum for Real-Valued Data: Real Time Estimation of the Panorama Through Circularity-Preserving Dft
  305. Complementary Set Variational Autoencoder for Supervised Anomaly Detection
  306. Complex Evolution Recurrent Neural Networks (ceRNNs)
  307. Complex-Valued Gaussian Process Latent Variable Model for Phase-Incorporating Speech Enhancement
  308. Complexity Reduction Algorithm for Optimum Quantizer Design Based on Amplitude Sparseness
  309. Complexity Reduction of Eigenvalue Decomposition-Based Diffuse Power Spectral Density Estimators Using the Power Method
  310. Compressed Convex Spectral Embedding for Bird Species Classification
  311. Compressed Sensing Mask Feature in Time-Frequency Domain for Civil Flight Radar Emitter Recognition
  312. Compressive Regularized Discriminant Analysis of High-Dimensional Data with Applications to Microarray Studies
  313. Compressive Sampling of Sound Fields Using Moving Microphones
  314. Compressive networked storage with lazy-encoding
  315. Computationally Efficient Iv-Based Bias Reduction for Closed-Form Tdoa Localization
  316. Concatenative Articulatory Video Synthesis Using Real-Time MRI Data for Spoken Language Training
  317. Concave Losses for Robust Dictionary Learning
  318. Concurrent Clutter and Noise Suppression via Low Rank Plus Sparse Optimization for Non-Contrast Ultrasound Flow Doppler Processing in Microvasculature
  319. Concurrent Target Following with Active Directional Sensors
  320. Confidence Based Acoustic Event Detection
  321. Confnet: Predict with Confidence
  322. Consecutive Independence and Correlation Transform for Multimodal Fusion: Application to Eeg and Fmri Data
  323. Considerations Regarding Individualization of Head-Related Transfer Functions
  324. Consistent Run Selection for Independent Component Analysis: Application to Fmri Analysis
  325. Constant False Alarm Rate for Online one Class Svm Learning
  326. Constant Modulus Probing Waveform Design for Mimo Radar Via Admm Algorithm
  327. Constrained Bayesian Active Learning of a Linear Classifier
  328. Constrained Convolutional-Recurrent Networks to Improve Speech Quality with Low Impact on Recognition Accuracy
  329. Content Delivery Design for Cache-Aided Cloud Radio Access Network to Achieve Low Latency
  330. Content-Based Representations of Audio Using Siamese Neural Networks
  331. Context-Sensitive Deep Learning for Detection of Clustered Micro Calcifications in Mammograms
  332. Continuous Security in IoT Using Blockchain
  333. Contourlet Based Natural Scene Statistics Using Student'S T Distribution
  334. Contrast Enhancement Using Phase Transition Information and Total Variation
  335. Control of Graph Signals Over Random Time-Varying Graphs
  336. Convergence Analysis on a Fast Iterative Phase Retrieval Algorithm Without Independence Assumption
  337. Convergence of Variance-Reduced Learning Under Random Reshuffling
  338. Convolutional Group-Sparse Coding and Source Localization
  339. Convolutional Neural Network Approach for Eeg-Based Emotion Recognition Using Brain Connectivity and its Spatial Information
  340. Convolutional Neural Networks and Multitask Strategies for Semantic Mapping of Natural Language Input to a Structured Database
  341. Convolutional Sequence to Sequence Model with Non-Sequential Greedy Decoding for Grapheme to Phoneme Conversion
  342. Convolutional Sparse Representations with Gradient Penalties
  343. Convolutional-Recurrent Neural Networks for Speech Enhancement
  344. Cooperative Tracking Using Marginal Diffusion Particle Filters
  345. Correlated Tensor Factorization for Audio Source Separation
  346. Correlation-Based Face Detection for Recognizing Faces in Videos
  347. Correntropy-Based Adaptive Filtering of Noncircular Complex Data
  348. Cortico-Muscular Coherence Enhancement Via Sparse Signal Representation
  349. Cover Song Identification Using Song-to-Song Cross-Similarity Matrix with Convolutional Neural Network
  350. Cramér-Rao Bound for Line Constrained Trajectory Tracking
  351. Crepe: A Convolutional Representation for Pitch Estimation
  352. Crime Incidents Embedding Using Restricted Boltzmann Machines
  353. Critically-Sampled Graph Filter Banks with Spectral Domain Sampling
  354. Cross-Lingual Phoneme Mapping for Language Robust Contextual Speech Recognition
  355. Cross-Modal Learning to Rank with Adaptive Listwise Constraint
  356. Cross-Modal Message Passing for Two-Stream Fusion
  357. Cross-Modality Distillation: A Case for Conditional Generative Adversarial Networks
  358. Cross-Validated Bandwidth Selection for Precision Matrix Estimation
  359. Crowdsourced Pairwise-Comparison for Source Separation Evaluation
  360. Crowdsourcing Emotional Speech
  361. Cyborg Speech: Deep Multilingual Speech Synthesis for Generating Segmental Foreign Accent with Natural Prosody
  362. DNN Based Embeddings for Language Recognition
  363. DNN Based Speaker Embedding Using Content Information for Text-Dependent Speaker Verification
  364. DNN-Based Concurrent Speakers Detector and its Application to Speaker Extraction with LCMV Beamforming
  365. Data Driven Convolutional Sparse Coding for Visual Recognition
  366. Data Injection Attack on Decentralized Optimization
  367. Data-Aided Fast Beamforming Selection for 5G
  368. Data-Driven Multi-Channel Filter Design with Peak-Interference Suppression for Threshold-Based Spike Sorting in High-Density Neural Probes
  369. Data-Driven Nonparametric Hypothesis Testing
  370. Decentralized Load Balancing in Mobile Communication Networks
  371. Deep Attractor Networks for Speaker Re-Identification and Blind Source Separation
  372. Deep Blind Image Quality Assessment by Learning Sensitivity Map
  373. Deep CNN Based Feature Extractor for Text-Prompted Speaker Recognition
  374. Deep Clustering with Gated Convolutional Networks
  375. Deep Factorization for Speech Signal
  376. Deep Feature Embedding Learning for Person Re-Identification Using Lifted Structured Loss
  377. Deep Feed-Forward Sequential Memory Networks for Speech Synthesis
  378. Deep Geometric Matrix Completion: A New Way for Recommender Systems
  379. Deep Image Super Resolution via Natural Image Priors
  380. Deep Layer Prior Optimization for Single Image Rain Streaks Removal
  381. Deep Learning Based Speech Beamforming
  382. Deep Learning for Accelerated Ultrasound Imaging
  383. Deep Learning for Frame Error Probability Prediction in BICM-OFDM Systems
  384. Deep Learning for Joint Source-Channel Coding of Text
  385. Deep Learning for Predicting Image Memorability
  386. Deep Mul Timodal Learning for Emotion Recognition in Spoken Language
  387. Deep Neural Network Based Discriminative Training for I-Vector/PLDA Speaker Verification
  388. Deep Residual Learning for Model-Based Iterative CT Reconstruction Using Plug-and-Play Framework
  389. Deep Residual Learning for Small-Footprint Keyword Spotting
  390. Deep Stock Representation Learning: From Candlestick Charts to Investment Decisions
  391. Deep Transfer Learning for EEG-Based Brain Computer Interface
  392. Deep Uniqueness-Aware Hashing for Fine-Grained Multi-Label Image Retrieval
  393. Deep Word Embeddings for Visual Speech Recognition
  394. Deep-FSMN for Large Vocabulary Continuous Speech Recognition
  395. Deepcasd: An End-to-End Approach for Multi-Spectral Image Super-Resolution
  396. Deeptongue: Tongue Segmentation Via Resnet
  397. Defending Against Packet-Size Side-Channel Attacks in Iot Networks
  398. Deformation Stability of Deep Convolutional Neural Networks on Sobolev Spaces
  399. Delta-Sigma Modulators for Constant Envelop Transmissions with Guaranteed Stability
  400. Demixing and Blind Deconvolution of Graph-Diffused Sparse Signals
  401. Demystifying Deep Learning: a Geometric Approach to Iterative Projections
  402. Densely Connected Progressive Learning for LSTM-Based Speech Enhancement
  403. Depression Speaks: Automatic Discrimination between Depressed and Non-Depressed Speakers Based on Nonverbal Speech Features
  404. Depth Super-Resolution Using Joint Adaptive Weighted Least Squares And Patching Gradient
  405. Depth Super-Resolution with Deep Edge-Inference Network and Edge-Guided Depth Filling
  406. Dereverberation and Beamforming in Far-Field Speaker Recognition
  407. Design of Optimal Entropy-Constrained Unrestricted Polar Quantizer for Bivariate Circularly Symmetric Sources
  408. Designing Signals with Good Correlation and Distribution Properties
  409. Detection of Cyclostationarity Using Generalized Coherence
  410. Determined Blind Source Separation via Proximal Splitting Algorithm
  411. Developing Far-Field Speaker System Via Teacher-Student Learning
  412. Developing a Geometric Deformable Model for Radar Shape Inversion
  413. Diabetic Retinopathy Detection Based on Deep Convolutional Neural Networks
  414. Dictionary Learning Algorithm for Multi-Subject Fmri Analysis Via Temporal and Spatial Concatenation
  415. Dictionary Learning for Gaussian Kernel Adaptive Filtering with Variablekernel Center and Width
  416. Dictionary Learning for High Dimensional Graph Signals
  417. Differentially Private Distributed Principal Component Analysis
  418. Digital-Analog Superposition Coding for Ofdm Channels with Application To Video Transmission
  419. Digitalseal: a Transaction Authentication Tool for Online and Offline Transactions
  420. Digraph Fourier Transform via Spectral Dispersion Minimization
  421. Direct Ensemble Estimation of Density Functionals
  422. Direct, Near Real Time Animation of a 3D Tongue Model Using Non-Invasive Ultrasound Images
  423. Directivity Synthesis with Multipoles Comprising a Cluster of Focused Sources Using a Linear Loudspeaker Array
  424. Directly Solving the Original Ratiocut Problem for Effective Data Clustering
  425. Discovering Correspondence Among Image Sets with Projection View Preservation For 3D Object Detection in Point Clouds
  426. Discriminative Clustering with Cardinality Constraints
  427. Discriminative Probabilistic Framework for Generalized Multi-Instance Learning
  428. Distributed Analytical Graph Identification
  429. Distributed Approximate Message Passing with Summation Propagation
  430. Distributed Censoring with Energy Constraint in Wireless Sensor Networks
  431. Distributed Coupled Learning Over Adaptive Networks
  432. Distributed Diffusion Adaptation Over Graph Signals
  433. Distributed Estimation Under Network Model Uncertainty
  434. Distributed Large Neural Network with Centralized Equivalence
  435. Distributed Maximum Likelihood Using Dynamic Average Consensus
  436. Distributed Model Construction in Radio Interferometric Calibration
  437. Distributed Optimal Consensus-Based Kalman Filtering and its Relation to Map Estimation
  438. Distributed Solution of Large-Scale Linear Systems Via Accelerated Projection-Based Consensus
  439. Distributed Splitting-Over-Features Sparse Bayesian Learning with Alternating Direction Method of Multipliers
  440. Distributed Submodular Maximization for Large Vocabulary Continuous Speech Recognition
  441. Distributed Tdoa-Based Indoor Source Localisation
  442. Dithered Beamforming for Channel Estimation in Mmwave-Based Massive Mimo
  443. Dnn-Based Ar-Wiener Filtering for Speech Enhancement
  444. Dnn-Based Voice Activity Detection Using Auxiliary Speech Models in Noisy Environments
  445. Dnn-Based Wireless Positioning in an Outdoor Environment
  446. Doa Estimation in Heteroscedastic Noise with Sparse Bayesian Learning
  447. Document Quality Estimation Using Spatial Frequency Response
  448. Domain Adversarial Training for Accented Speech Recognition
  449. Domain Independent Key Term Extraction from Spoken Content Based on Context and Term Location Information in the Utterances
  450. Domain and Speaker Adaptation for Cortana Speech Recognition
  451. Dpca: Dimensionality Reduction for Discriminative Analytics of Multiple Large-Scale Datasets
  452. Driver Estimation in Non-Linear Autoregressive Models
  453. Dropout Approaches for LSTM Based Speech Recognition Systems
  454. Dual Frequency- and Block-Permutation Alignment for Deep Learning Based Block-Online Blind Source Separation
  455. Dual-Channel Modulation Energy Metric for Direct-to-Reverberation Ratio Estimation
  456. Dynamic Frame Skipping for Fast Speech Recognition in Recurrent Neural Network Based Acoustic Models
  457. Dynamic Matrix Recovery from Partially Observed and Erroneous Measurements
  458. Dynamic Multi-Rater Gaussian Mixture Regression Incorporating Temporal Dependencies of Emotion Uncertainty Using Kalman Filters
  459. EAR-EEG for Detecting Inter-Brain Synchronisation in Continuous Cooperative Multi-Person Scenarios
  460. EEG-Based Auditory Attention Decoding Using Steerable Binaural Superdirective Beamformer
  461. Eadnet: Efficient Architecture for Decomposed Convolutional Neural Networks
  462. Ecg Delineation for Qt Interval Analysis Using an Unsupervised Learning Method
  463. Edge-Aware Context Encoder for Image Inpainting
  464. Edge-Based Loss Function for Single Image Super-Resolution
  465. Eeg-Based Video Identification Using Graph Signal Modeling and Graph Convolutional Neural Network
  466. Effective Attention Mechanism in Dynamic Models for Speech Emotion Recognition
  467. Effective Cover Song Identification Based on Skipping Bigrams
  468. Effective Noise Removal and Unified Model of Hybrid Feature Space Optimization for Automated Cardiac Anomaly Detection Using Phonocardiogarm Signals
  469. Efficacy of Multiuser Massive Miso Wireless Energy Transfer Under iq Imbalance and Channel Estimation Errors Over Rician Fading
  470. Efficient Circulant Matrix Construction and Implementation for Compressed Sensing
  471. Efficient Constrained Tensor Factorization by Alternating Optimization with Primal-Dual Splitting
  472. Efficient Convolutional Dictionary Learning Using Partial Update Fast Iterative Shrinkage-Thresholding Algorithm
  473. Efficient Deep Convolutional Neural Networks Accelerator without Multiplication and Retraining
  474. Efficient Estimation of Scatter Matrix with Convex Structure Under $T$ -Distribution
  475. Efficient Integration of Fixed Beamformers and Speech Separation Networks for Multi-Channel Far-Field Speech Separation
  476. Efficient Learning of Articulatory Models Based on Multi-Label Training and Label Correction for Pronunciation Learning
  477. Efficient Model-Free Learning to Overcome Hardware Nonidealities in Analog-to-Information Converters
  478. Efficient Non-Convex Graph Clustering for Big Data
  479. Efficient Sampling on HEALPix Grid
  480. Efficient Super-Wide Bandwidth Extension Using Linear Prediction Based Analysis-Synthesis
  481. Efficient Worker Assignment in Crowdsourced Data Labeling Using Graph Signal Processing
  482. Efficiently Trainable Text-to-Speech System Based on Deep Convolutional Networks with Guided Attention
  483. Em-Based Semi-Blind Mimo-Ofdm Channel Estimation
  484. Emg Acquisition and Hand Pose Classification for Bionic Hands from Randomly-Placed Sensors
  485. Emphatic Speech Generation with Conditioned Input Layer and Bidirectional LSTMS for Expressive Speech Synthesis
  486. Emphatic Speech Prosody Prediction with Deep Lstm Networks
  487. Enabling Early Audio Event Detection with Neural Networks
  488. End-To-End Low-Resource Lip-Reading with Maxout Cnn and Lstm
  489. End-To-End Optimized Speech Coding with Deep Neural Networks
  490. End-to-End Audiovisual Speech Recognition
  491. End-to-End Automatic Speech Translation of Audiobooks
  492. End-to-End Continuous Emotion Recognition from Video Using 3D Convlstm Networks
  493. End-to-End DNN Based Speaker Recognition Inspired by I-Vector and PLDA
  494. End-to-End Dynamic Query Memory Network for Entity-Value Independent Task-Oriented Dialog
  495. End-to-End Hierarchical Language Identification System
  496. End-to-End Multi-Speaker Speech Recognition
  497. End-to-End Neural Network Based Automated Speech Scoring
  498. End-to-End Sound Source Enhancement Using Deep Neural Network in the Modified Discrete Cosine Transform Domain
  499. End-to-End Speech Emotion Recognition Using Deep Neural Networks
  500. End-to-end Multimodal Speech Recognition
  501. Endmembers as Directional Data for Robust Material Variability Retrieval in Hyperspectral Image Unmixing
  502. Energy Efficiency in MIMO Interference Channels: Social Optimality and Max-Min Fairness
  503. Energy Efficient Consensus Over Directed Graphs
  504. Energy-Efficient Speaker Identification with Low-Precision Networks
  505. Enhancement and Analysis of Conversational Speech: JSALT 2017
  506. Entropy Based Pruning of Backoff Maxent Language Models with Contextual Features
  507. Envelope Estimation by Tangentially Constrained Spline
  508. Epileptic State Segmentation with Temporal-Constrained Clustering
  509. Essence Vector-Based Query Modeling for Spoken Document Retrieval
  510. Estimation of Source Panning Parameters and Segmentation of Stereophonic Mixtures
  511. Estimation of Time-Varying Room Impulse Responses of Multiple Sound Sources from Observed Mixture and Isolated Source Signals
  512. Estimation of the Sound Field at Arbitrary Positions in Distributed Microphone Networks Based on Distributed Ray Space Transform
  513. Evaluating Models of Dynamic Functional Connectivity Using Predictive Classification Accuracy
  514. Evaluation of the Penalized Inequality Constrained Minimum Variance Beamformer for Hearing Aids
  515. Event-Triggered Particle Filtering Via Diffusion Strategies for Distributed Estimation in Autonomous Systems
  516. Eventness: Object Detection on Spectrograms for Temporal Localization of Audio Events
  517. Evolutionary Spectra Based on the Multitaper Method with Application To Stationarity Test
  518. Exploitation of Semantic Keywords for Malicious Event Classification
  519. Exploiting Convolutional Neural Networks for Phonotactic Based Dialect Identification
  520. Exploiting Explicit Memory Inclusion for Artificial Bandwidth Extension
  521. Exploring Ctc-Network Derived Features with Conventional Hybrid System
  522. Exploring Hashing and Cryptonet Based Approaches for Privacy-Preserving Speech Emotion Recognition
  523. Exploring Motor Imagery Eeg Patterns for Stroke Patients with Deep Neural Networks
  524. Exploring Practical Aspects of Neural Mask-Based Beamforming for Far-Field Speech Recognition
  525. Exploring Sequential Characteristics in Speaker Bottleneck Feature for Text-Dependent Speaker Verification
  526. Exploring Speech Enhancement with Generative Adversarial Networks for Robust Speech Recognition
  527. Exploring the Non-Local Similarity Present in Variational Mode Functions for Effective ECG Denoising
  528. Exploring the Use of Group Delay for Generalised VTS Based Noise Compensation
  529. Exponentially Consistent K-Means Clustering Algorithm Based on Kolmogrov-Smirnov Test
  530. Extendable Neural Matrix Completion
  531. Extended Pipeline for Content-Based Feature Engineering in Music Genre Recognition
  532. Extension and Evaluation of a Spectro-Temporal Modulation Method to Improve Acoustic Feedback Performance in Hearing Aids
  533. Extension of Decoding Problem of HMM Based on LP-Norm
  534. Extracting Domain Invariant Features by Unsupervised Learning for Robust Automatic Speech Recognition
  535. F0 Estimation for DNN-Based Ultrasound Silent Speech Interfaces
  536. FDD Massive MIMO Channel Spatial Covariance Conversion Using Projection Methods
  537. Face Hallucination Based on Key Parts Enhancement
  538. Facial Feature-Integrated Inter-Camera Human Tracking
  539. Factorized Hidden Variability Learning for Adaptation of Short Duration Language Identification Models
  540. Fairness in Multiterminal Data Compression: A Splitting Method for the Egalitarian Solution
  541. Far-Field Audio-Visual Scene Perception of Multi-Party Human-Robot Interaction for Children and Adults
  542. Fast 3D-Hevc Depth Maps Intra-Frame Prediction Using Data Mining
  543. Fast Adaptation on Deepmixture Generative Network Based Acoustic Modeling
  544. Fast And Robust Recursive Filter for Image Denoising
  545. Fast Decentralized Learning Via Hybrid Consensus Admm
  546. Fast Detection of Abnormal Events in Videos with Binary Features
  547. Fast Dictionary-Based Approach for Mass Spectrometry Data Analysis
  548. Fast Distributed Subspace Projection via Graph Filters
  549. Fast Oov Words Incorporation Using Structured Word Embeddings for Neural Network Language Model
  550. Fast Projection onto the 𝓁∞, 1-Mixed Norm Ball Using Steffensen Root Search
  551. Fast Projection-Based Solvers for the Non-Convex Quadratically Constrained Feasibility Problem
  552. Fast Robust Tracking Via Double Correlation Filter Formulation
  553. Fast Texture Intra Size Coding Based On Big Data Clustering for 3D-Hevc
  554. Fast Variational Level Set Based Image Segmentation via Two-Scale Filtering Model
  555. Fast Vehicle Detection with Lateral Convolutional Neural Network
  556. Fast and Adaptive Blind Audio Source Separation Using Recursive Levenberg-Marquardt Synchrosqueezing
  557. Fast-Convergence Singular Value Decomposition for Tracking Time-Varying Channels in Massive Mimo Systems
  558. Faster ICA Under Orthogonal Constraint
  559. Faster and Still Safe: Combining Screening Techniques and Structured Dictionaries to Accelerate the Lasso
  560. Faster-Than-Nyquist Signaling with Differential Encoding and Non Coherent Detection
  561. Fault Detection Using Attention Models Based on Visual Saliency
  562. Feature Based Adaptation for Speaking Style Synthesis
  563. Feature Design Using Audio Decomposition for Intelligent Control of the Dynamic Range Compressor
  564. Feature LMS Algorithms
  565. Feature Matching Based on Top K Rank Similarity
  566. Fftnet: A Real-Time Speaker-Dependent Neural Vocoder
  567. Filter-and-Convolve: A Cnn Based Multichannel Complex Concatenation Acoustic Model
  568. Fine-Grained Wound Tissue Analysis Using Deep Neural Network
  569. Finite Sample Performance of Linear Least Squares Estimators Under Sub-Gaussian Martingale Difference Noise
  570. Finite-Alphabet Noma for Two-User Uplink Channel
  571. First-Order Bifurcation Detection for Dynamic Complex Networks
  572. First-Order Difference Energy Regularization for Enhancing Reconstruction Performance in Compressive Sensing of Foot-Gait Signals
  573. First-Order Perturbation Analysis of Secsi With Generalized Unfoldings
  574. Flexible Multi-Group Single-Carrier Modulation: Optimal Subcarrier Grouping and Rate Maximization
  575. Flipping Large Classes on a Shoestring Budget
  576. Focal Kl-Divergence Based Dilated Convolutional Neural Networks for Co-Channel Speaker Identification
  577. Fooling End-To-End Speaker Verification With Adversarial Examples
  578. Foreground Harmonic Noise Reduction for Robust Audio Fingerprinting
  579. Forward Attention in Sequence- To-Sequence Acoustic Modeling for Speech Synthesis
  580. Forward Vehicle Collision Warning Based on Quick Camera Calibration
  581. Fps-Sft: A Multi-Dimensional Sparse Fourier Transform Based on the Fourier Projection-Slice Theorem
  582. Frame-Subsampled, Drift-Resilient Video Object Tracking
  583. Frame-by-Frame Closed-Form Update for Mask-Based Adaptive MVDR Beamforming
  584. Framework for Evaluation of Sound Event Detection in Web Videos
  585. Frontal Face Generation from Multiple Pose-Variant Faces with CGAN in Real-World Surveillance Scene
  586. Full-Info Training for Deep Speaker Feature Learning
  587. Full-Reference Quality Assessment of Contrast Changed Images Based on Local Linear Model
  588. Fully Automatic Segmentation of the Right Ventricle Via Multi-Task Deep Neural Networks
  589. Functional Connectivity States of the Brain Using Restricted Boltzmann Machines
  590. Fusion of Multiple Multiband Images with Complementary Spatial and Spectral Resolutions
  591. GLRT Particle Filter for Tracking Nlos Target in Around-the-Corner Radar
  592. GM-PHD Filter Based Online Multiple Human Tracking Using Deep Discriminative Correlation Matching
  593. GMM-Based Iterative Entropy Coding for Spectral Envelopes of Speech and Audio
  594. GSC-Based Binaural Speaker Separation Preserving Spatial Cues
  595. Gamification of DSP: Electronic vs Pen-and-Paper
  596. Gated Residual Networks with Dilated Convolutions for Supervised Speech Separation
  597. Generalised Discriminative Transform via Curriculum Learning for Speaker Recognition
  598. Generalised Sidelobe Canceller for Noise Reduction in Hearing Devices Using an External Microphone
  599. Generalization of Deep Neural Networks for Chest Pathology Classification in X-Rays Using Generative Adversarial Networks
  600. Generalized End-to-End Loss for Speaker Verification
  601. Generalized Linear Mixing Model Accounting for Endmember Variability
  602. Generalized Tensor Contractions for an Improved Receiver Design in MIMO-OFDM Systems
  603. Generalized Uncertainty Principles for the Two-Sided Quaternion Linear Canonical Transform
  604. Generating Sound Words from Audio Signals of Acoustic Events with Sequence-to-Sequence Model
  605. Generative Adversarial Networks Based Data Augmentation for Noise Robust Speech Recognition
  606. Generative Adversarial Source Separation
  607. Generative Model and Associated Metric for Coordinated-Motion Target Groups
  608. Generative Scatternet Hybrid Deep Learning (G-Shdl) Network with Structural Priors for Semantic Image Segmentation
  609. Geographic Language Models for Automatic Speech Recognition
  610. Geolocation of Unknown Emitters Using Tdoa of Path Rays Through the Ionosphere by Multiple Coordinated Distant Receivers
  611. Geometric Information Based Monaural Speech Separation Using Deep Neural Network
  612. Geometric Transformation Invariant Image Quality Assessment Using Convolutional Neural Networks
  613. Global Optimality in Inductive Matrix Completion
  614. Globally Optimal Energy Efficiency Maximization for Capacity-Limited Fronthaul Crans with Dynamic Power Amplifiers' Efficiency
  615. Graph Error Effect in Graph Signal Processing
  616. Graph Learning Based on Total Variation Minimization
  617. Graph Regularized Tensor Factorization for Single-Trial EEG Analysis
  618. Graph Sampling with and Without Input Priors
  619. Graph Signal Processing of Human Brain Imaging Data
  620. Graph-based Transforms for Predictive Light Field Compression based on Super-Pixels
  621. Grassmann Singular Spectrum Analysis for Bioacoustics Classification
  622. Greedy Algorithm with Approximation Ratio for Sampling Noisy Graph Signals
  623. Greedy Pursuits Based Gradual Weighting Strategy for Weighted $\ell_{1}$-Minimization
  624. Grid-Free Direction-of-Arrival Estimation with Compressed Sensing and Arbitrary Antenna Arrays
  625. Gridless Two-Dimensional Doa Estimation With L-Shaped Array Based on the Cross-Covariance Matrix
  626. Group Sparsity Residual with Non-Local Samples for Image Denoising
  627. Guided Image Filtering with Arbitrary Window Function
  628. HNSR: Highway Networks Based Deep Convolutional Neural Networks Model for Single Image Super-Resolution
  629. Hand-Raising Gesture Detection in Real Classroom
  630. Hand: Header-Assisted Network Decoding
  631. Hands-on in Signal Processing Education at Technische Universitat Darmstadt
  632. Hard Shadows Removal Using an Approximate Illumination Invariant
  633. Harnessing Bandit Online Learning to Low-Latency Fog Computing
  634. Hi, Bcd! Hybrid Inexact Block Coordinate Descent for Hyperspectral Super-Resolution
  635. Hierarchical Attention and Context Modeling for Group Activity Recognition
  636. Hierarchical Heavy Hitter Detection Under Unknown Models
  637. Hierarchical Segmentation Based Point Cloud Attribute Compression
  638. High Accuracy Acoustic Estimation of Multiple Targets
  639. High Efficiency Compression for Object Detection
  640. High Order Recurrent Neural Networks for Acoustic Modelling
  641. High-Accuracy Stochastic Computing-Based FIR Filter Design
  642. High-Order Tensor Completion for Data Recovery via Sparse Tensor-Train Optimization
  643. High-Quality Nonparallel Voice Conversion Based on Cycle-Consistent Adversarial Network
  644. High-Speed Light Field Image Formation Analysis Using Wavefield Modeling with Flexible Sampling
  645. High-Speed Optical Camera Communication Using an Optimally Modulated Signal
  646. Higher Order Exponential Splittings for the Fast Non-Linear Fourier Transform of the Korteweg-De Vries Equation
  647. Hough Transform Guided Deep Feature Extraction for Dense Building Detection in Remote Sensing Images
  648. How Sampling Rate Affects Cross-Domain Transfer Learning for Video Description
  649. How are the Centered Kernel Principal Components Relevant to Regression Task? -An Exact Analysis
  650. How to Interconnect for Massive Mimo Self-Calibration?
  651. How to Mobilize Mmwave: A Joint Beam and Channel Tracking Approach
  652. Human Motion Classification with Micro-Doppler Radar and Bayesian-Optimized Convolutional Neural Networks
  653. Human and Machine Speaker Recognition Based on Short Trivial Events
  654. Human and Machine Type Communications Can Coexist in Uplink Massive Mimo Systems
  655. Human-Like Emotion Recognition: Multi-Label Learning from Noisy Labeled Audio-Visual Expressive Speech
  656. Human-Machine Inference Networks for Smart Decision Making: Opportunities and Challenges
  657. Hybrid Lstm-Fsmn Networks for Acoustic Modeling
  658. Hybridnet for Depth Estimation and Semantic Segmentation
  659. Hypercomplex Tensor Completion with Cayley-Dickson Singular Value Decomposition
  660. Hyperspectral Super-Resolution Via Coupled Tensor Factorization: Identifiability and Algorithms
  661. ILAPF: Incremental Learning Assisted Particle Filtering
  662. IVA-Based Spatio-Temporal Dynamic Connectivity Analysis in Large-Scale FMRI Data
  663. Identification of Bilinear Forms with the Kalman Filter
  664. Identification of Multiple-Input Multiple-Output Channels Under Linear Side Constraints
  665. Identifying Susceptible Agents in Time Varying Opinion Dynamics Through Compressive Measurements
  666. Identifying Undirected Network Structure via Semidefinite Relaxation
  667. Image Alignment via Multi-Model Geometric Fitting and Hierarchical Homography Estimation
  668. Image Augmentation Using Radial Transform for Training Deep Neural Networks
  669. Image Fusion Using Belief Propagation
  670. Image Fusion: an Introduction to Multispectral Signal Processing
  671. Image Quality Assessment Based Label Smoothing in Deep Neural Network Learning
  672. Image Recognition Based on Separable Lattice Hmms Using a Deep Neural Network for Output Probability Distributions
  673. Image Reconstruction for Quanta Image Sensors Using Deep Neural Networks
  674. Image Representation Using Supervised and Unsupervised Learning Methods on Complex Domain
  675. Image Restoration with Deep Generative Models
  676. Image-Based PM2.5 Estimation and its Application on Depth Estimation
  677. Impact of Microphone Array Configurations on Robust Indirect 3d Acoustic Source Localization
  678. Importance Sampling Estimator of Outage Probability under Generalized Selection Combining Model
  679. Improved Algorithms for Differentially Private Orthogonal Tensor Decomposition
  680. Improved Audio-Visual Laughter Detection Via Multi-Scale Multi-Resolution Image Texture Features and Classifier Fusion
  681. Improved Detection of Semi-Percussive Onsets in Audio Using Temporal Reassignment
  682. Improved Noise Characterization for Relative Impulse Response Estimation
  683. Improved Steady State Analysis of the Recursive Least Squares Algorithm
  684. Improved Tdnns Using Deep Kernels and Frequency Dependent Grid-RNNS
  685. Improved Weighted Instrumental Variable Estimator for Doppler-Bearing Source Localization in Heavy Noise
  686. Improving Accuracy of Nonparametric Transfer Learning Via Vector Segmentation
  687. Improving Consensus-Based Distributed Camera Calibration Via Edge Pruning and Graph Traversal Initialization
  688. Improving Convolutional Neural Networks Via Compacting Features
  689. Improving Disparity Map Estimation for Multi-View Noisy Images
  690. Improving End-of-Turn Detection in Spoken Dialogues by Detecting Speaker Intentions as a Secondary Task
  691. Improving End-to-End Speech Recognition with Policy Learning
  692. Improving Mandarin Tone Mispronunciation Detection for Non-Native Learners with Soft-Target Tone Labels and BLSTM-Based Deep Models
  693. Improving Multichannel Speech Recognition with Generalized Cross Correlation Inputs and Multitask Learning
  694. Improving Multikernel Adaptive Filtering with Selective Bias
  695. Improving Sar Automatic Target Recognition Using Simulated Images Under Deep Residual Refinements
  696. Improving Semi-Supervised Classification for Low-Resource Speech Interaction Applications
  697. Improving the Capacity of Very Deep Networks with Maxout Units
  698. Improving the Performance of Online Neural Transducer Models
  699. Incorporating ASR Errors with Attention-Based, Jointly Trained RNN for Intent Detection and Slot Filling
  700. Incorporating Scalability in Unsupervised Spatio- Temporal Feature Learning
  701. Independent Low-Rank Matrix Analysis Based on Multivariate Complex Exponential Power Distribution
  702. Indian Buffet Process Deep Generative Models for Semi-Supervised Classification
  703. Individual Difference of Ultrasonic Transducers for Parametric Array Loudspeaker
  704. Individual Ship Detection Using Underwater Acoustics
  705. Inexact Proximal Operators for 𝓁p-Quasinorm Minimization
  706. Influence of the Number of Loudspeakers on the Timbre in Mixed-Order Ambisonics Reprodution
  707. Information Fusion Using Particles Intersection
  708. Insense: Incoherent Sensor Selection for Sparse Signals
  709. Insights in-to-End Learning Scheme for Language Identification
  710. Instlistener: An Expressive Parameter Estimation System Imitating Human Performances of Monophonic Musical Instruments
  711. Integrating Perceivers Neural-Perceptual Responses Using a Deep Voting Fusion Network for Automatic Vocal Emotion Decoding
  712. Intelligent Signal Processing Mechanisms for Nuanced Anomaly Detection in Action Audio-Visual Data Streams
  713. Interference Reduction on Full-Length Live Recordings
  714. Interpretable Clustering Ensembles Using Binary Matrix Factorization
  715. Interpreting DNN Output Layer Activations: A Strategy to Cope with Unseen Data in Speech Recognition
  716. Invariances and Data Augmentation for Supervised Music Transcription
  717. Inverse Atmoshperic Scattering Modeling with Convolutional Neural Networks for Single Image Dehazing
  718. Investigating Label Noise Sensitivity of Convolutional Neural Networks for Fine Grained Audio Signal Labelling
  719. Investigating the Effect of Sound-Event Loudness on Crowdsourced Audio Annotations
  720. Investigation in Spatial-Temporal Domain for Face Spoof Detection
  721. Investigations on End- to-End Audiovisual Fusion
  722. Invisible Geo-Location Signature in A Single Image
  723. Iterative Deep Neural Networks for Speaker-Independent Binaural Blind Speech Separation
  724. JND-Based Perceptual Video Coding for 4: 4: 4 Screen Content Data in HEVC
  725. Joint Adaptive Impulse Response Estimation and Inverse Filtering for Enhancing In-Car Audio
  726. Joint Audio-Video Driven Facial Animation
  727. Joint Estimation of the Room Geometry and Modes with Compressed Sensing
  728. Joint Gender-, Tone-, Vowel- Classification Via Novel Hierarchical Classification for Annotation of Monosyllabic Mandarin Word Tokens
  729. Joint I-Vector with End-to-End System for Short Duration Text-Independent Speaker Verification
  730. Joint Independent Subspace Analysis by Coupled Block Decomposition: Non-Identifiable Cases
  731. Joint Late Reverberation and Noise Power Spectral Density Estimation in a Spatially Homogeneous Noise Field
  732. Joint License Plate Super-Resolution and Recognition in One Multi-Task Gan Framework
  733. Joint List Polar Decoder with Successive Cancellation and Sphere Decoding
  734. Joint Mobile Sink Scheduling and Data Aggregation in Asynchronous Wireless Sensor Networks Using Q-Learning
  735. Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition
  736. Joint Probabilistic Forecasts of Temperature and Solar Irradiance
  737. Joint Screening Tests for Lasso
  738. Joint Separation and Dereverberation of Reverberant Mixtures with Determined Multichannel Non-Negative Matrix Factorization
  739. Joint Source Localization and Dereverberation by Sound Field Interpolation Using Sparse Regularization
  740. Joint Source and Sensor Placement for Sound Field Control Based on Empirical Interpolation Method
  741. Joint Space-(Slow) Time Transmission with Unimodular Waveforms and Receive Adaptive Filter Design for Radar
  742. Joint Speaker Diarization and Recognition Using Convolutional and Recurrent Neural Networks
  743. Joint Time Synchronization and Localization for Target Sensors Using a Single Mobile Anchor with Position Uncertainties
  744. Joint Topology and Radio Resource Optimization for Device-to-Device Based Mobile Social Networks
  745. Joint Verification-Identification in end-to-end Multi-Scale CNN Framework for Topic Identification
  746. Jointly Tracking and Separating Speech Sources Using Multiple Features and the Generalized Labeled Multi-Bernoulli Framework
  747. Kalman Filtering and Clustering in Sensor Networks
  748. Kernel-Induced Sampling Theorem for Translation-Invariant Reproducing Kernel Hilbert Spaces with Uniform Sampling
  749. Knowledge Transfer from Weakly Labeled Audio Using Convolutional Neural Network for Sound Events and Scenes
  750. Knowledge Transfer in Permutation Invariant Training for Single-Channel Multi-Talker Speech Recognition
  751. L1 Patch-Based Image Partitioning into Homogeneous Textured Regions
  752. Language Model Domain Adaptation Via Recurrent Neural Networks with Domain-Shared and Domain-Specific Representations
  753. Language Transfer of Audio Word2Vec: Learning Audio Segment Representations Without Target Language Data
  754. Language and Noise Transfer in Speech Enhancement Generative Adversarial Network
  755. Large-Scale High-Dimensional Clustering with Fast Sketching
  756. Large-Scale Regularized Sumcor GCCA via Penalty-Dual Decomposition
  757. Large-Scale Weakly Supervised Audio Classification Using Gated Convolutional Neural Network
  758. Late Reverberation Suppression Using Recurrent Neural Networks with Long Short-Term Memory
  759. Learned Convolutional Sparse Coding
  760. Learned Forensic Source Similarity for Unknown Camera Models
  761. Learning Deep Representations Using Convolutional Auto-Encoders with Symmetric Skip Connections
  762. Learning Explicit Shape and Motion Evolution Maps for Skeleton-Based Human Action Recognition
  763. Learning Filterbanks from Raw Speech for Phone Recognition
  764. Learning Gaussian Graphical Models Using Discriminated Hub Graphical Lasso
  765. Learning Hard Alignments with Variational Inference
  766. Learning In-Place Residual Homogeneity for Image Detail Enhancement
  767. Learning Lexical Coherence Representation Using LSTM Forget Gate for Children with Autism Spectrum Disorder During Story-Telling
  768. Learning Neural Trans-Dimensional Random Field Language Models with Noise-Contrastive Estimation
  769. Learning Statistically Accurate Resource Allocations in Non-Stationary Wireless Systems
  770. Learning Temporal Relationships Between Financial Signals
  771. Learning an Inverse Tone Mapping Network with a Generative Adversarial Regularizer
  772. Learning on a Budget for User Authentication on Mobile Devices
  773. Learning-Based Acoustic Source-Microphone Distance Estimation Using the Coherent-to-Diffuse Power Ratio
  774. Learning-Based Complexity Reduction and Scaling for HEVC Encoders
  775. Learning-Based Design of Measurement Matrix with Inter-Column Correlation for Compressive Sensing
  776. Lensless 3D Imaging Using Mask-Based Cameras
  777. Leveraging LSTM Models for Overlap Detection in Multi-Party Meetings
  778. Lexico-Acoustic Neural-Based Models for Dialog Act Classification
  779. Limited-Memory BFGS Optimization of Recurrent Neural Network Language Models for Speech Recognition
  780. Limiting Numerical Precision of Neural Networks to Achieve Real-Time Voice Activity Detection
  781. Linear Classification in Speech-Based Objective Differential Diagnosis of Parkinsonism
  782. Linear Networks Based Speaker Adaptation for Speech Synthesis
  783. Linear Quantization by Effective-Resistance Sampling
  784. Linguistic Unit Discovery from Multi-Modal Inputs in Unwritten Languages: Summary of the "Speaking Rosetta" JSALT 2017 Workshop
  785. Lip2Audspec: Speech Reconstruction from Silent Lip Movements Video
  786. Listening to Each Speaker One by One with Recurrent Selective Hearing Networks
  787. Lo-Regularized Hybrid Gradient Sparsity Priors for Robust Single-Image Blind Deblurring
  788. Locality-Preserving Complex-Valued Gaussian Process Latent Variable Model for Robust Face Recognition
  789. Localization-Free Power Cartography
  790. Locally Optimal Invariant Detector for Testing Equality of Two Power Spectral Densities
  791. Loudspeaker and Listening Position Estimation Using Smart Speakers
  792. Low Complexity Heart Rate Measurement from Wearable Wrist-Type Photoplethysmographic Sensors Robust to Motion Artifacts
  793. Low Complexity Implementation of Carrier and Symbol Timing Synchronization for a Fully Digital Downhole Telemetry System
  794. Low Complexity Joint RDO of Prediction Units Couples for HEVC Intra Coding
  795. Low Rank Fourier Ptychography
  796. Low Resolution Face Recognition and Reconstruction Via Deep Canonical Correlation Analysis
  797. Low-Complexity Secure Watermark Encryption for Compressed Sensing-Based Privacy Preserving
  798. Low-Complexity Weighted Mrt Multicast Beamforming in Massive Mimo Cellular Networks
  799. Low-Energy Graph Fourier Basis Functions Span Salient Objects
  800. Low-Overhead Receiver-Side Channel Tracking for Mmwave Mimo
  801. Low-Rank Matrix Recovery from One-Bit Comparison Information
  802. Low-Rank Optimization for Data Shuffling in Wireless Distributed Computing
  803. Low-Rank and Joint-Sparse Signal Recovery for Spatially and Temporally Correlated Data Using Sparse Bayesian Learning
  804. MIMO Radar Target Detection Using Low-Complexity Receiver
  805. MIMO Transmit Beampattern Matching Under Waveform Constraints
  806. MMSE Adaptive Waveform Design for a MIMO Active Sensing System Tracking Multiple Moving Targets
  807. Machine Assisted Human Decision Making
  808. Machine Load Estimation Via Stacked Autoencoder Regression
  809. Man-Made Object Recognition from Underwater Optical Images Using Deep Learning and Transfer Learning
  810. Manifold-Based Analysis of Natural Stochastic Textures with Application in Texture Synthesis
  811. Manifold-Based Inference for a Supervised Gaussian Process Classifier
  812. Marginal Bayesian Bhattacharyya Bounds for Discrete-Time Filtering
  813. Mask Weighted Stft Ratios for Relative Transfer Function Estimation and ITS Application to Robust ASR
  814. Matching Projection Decoding Method for Ambisonics System
  815. Matching Pursuit Based Convolutional Sparse Coding
  816. Matrix Completion as Graph Bandlimited Reconstruction
  817. Maximal Figure-of-Merit Embedding for Multi-Label Audio Classification
  818. Maximum-A-Posteriori Signal Recovery with Prior Information: Applications to Compressive Sensing
  819. Maximum-Likelihood Online Speaker Diarization in Noisy Meetings Based on Categorical Mixture Model and Probabilistic Spatial Dictionary
  820. Measuring Uncertainty in Deep Regression Models: The Case of Age Estimation from Speech
  821. Measuring the Effect of Linguistic Resources on Prosody Modeling for Speech Synthesis
  822. Meeting Recognition with Asynchronous Distributed Microphone Array Using Block-Wise Refinement of Mask-Based MVDR Beamformer
  823. Method of Estimating Direction of Arrival of Sound Source for Monaural Hearing Based on Temporal Modulation Perception
  824. Mgn: Multi-Glimpse Network for Action Recognition
  825. Min-Max Latency Optimization for Multiuser Computation Offloading in Fog-Radio Access Networks
  826. Minimum Spanning Distance for Image Segmentation
  827. Minimum Word Error Rate Training for Attention-Based Sequence-to-Sequence Models
  828. Mitigation of Nonlinear Distortion in Sound Zone Control by Constraining Individual Loudspeaker Driver Amplitudes
  829. Mmse-Based Autocorrelation Sampling for Comprime Arrays
  830. Mobile Bayesian Spectrum Learning for Heterogeneous Networks
  831. Modal Decomposition of Musical Instrument Sound Via Alternating Direction Method of Multipliers
  832. Modality-Specific Structure Preserving Hashing for Cross-Modal Retrieval
  833. Mode Domain Spatial Active Noise Control Using Sparse Signal Representation
  834. Model-Based Free-Breathing Cardiac MRI Reconstruction Using Deep Learned & Storm Priors: MODL-STORM
  835. Model-Based Noise PSD Estimation from Speech in Non-Stationary Noise
  836. Modeling Non-Linguistic Contextual Signals in LSTM Language Models Via Domain Adaptation
  837. Modeling and Detection of Evolving Threats Using Random Finite Set Statistics
  838. Modeling the Acquisition of Intonation: A First Step
  839. Modeling-By-Generation-Structured Noise Compensation Algorithm for Glottal Vocoding Speech Synthesis System
  840. Modelling Jitter in Wireless Channel Created by Processor-Memory Activity
  841. Monaural Singing Voice Separation with Skip-Filtering Connections and Recurrent Inference of Time-Frequency Mask
  842. Monaural Speech Enhancement Using Deep Neural Networks by Maximizing a Short-Time Objective Intelligibility Measure
  843. Monophone-Based Background Modeling for Two-Stage On-Device Wake Word Detection
  844. Motor Imagery for Eeg Biometrics Using Convolutional Neural Network
  845. Mrt-Based Joint Unicast and Multigroup Multicast Transmission in Massive Mimo Systems
  846. Mse-Optimal 1-Bit Precoding for Multiuser Mimo Via Branch and Bound
  847. Multi Scale Feedback Connection for Noise Robust Acoustic Modeling
  848. Multi Task Learning with Positive and Unlabeled Data and its Application to Mental State Prediction
  849. Multi-Armed Bandits for Human-Machine Decision Making
  850. Multi-Channel Deep Clustering: Discriminative Spectral and Spatial Embeddings for Speaker-Independent Speech Separation
  851. Multi-Dialect Speech Recognition with a Single Sequence-to-Sequence Model
  852. Multi-Exposure Image Fusion Based on Exposure Compensation
  853. Multi-Kernel Regression for Graph Signal Processing
  854. Multi-Kernel, Deep Neural Network and Hybrid Models for Privacy Preserving Machine Learning
  855. Multi-Microphone Neural Speech Separation for Far-Field Multi-Talker Speech Recognition
  856. Multi-Scale Object Detection with Feature Fusion and Region Objectness Network
  857. Multi-Scale Recurrent Neural Network for Sound Event Detection
  858. Multi-Scenario Deep Learning for Multi-Speaker Source Separation
  859. Multi-Segment Reconstruction Using Invariant Features
  860. Multi-Task Autoencoder for Noise-Robust Speech Recognition
  861. Multi-View Audio-Articulatory Features for Phonetic Recognition on RTMRI-TIMIT Database
  862. Multi-View Source Localization Based on Power Ratios
  863. Multichannel Kalman Filtering for Speech Ehnancement
  864. Multichannel Speaker Activity Detection for Meetings
  865. Multichannel Speech Separation with Recurrent Neural Networks from High-Order Ambisonics Recordings
  866. Multilayer Adaptation Based Complex Echo Cancellation and Voice Enhancement
  867. Multilingual Adaptation of RNN Based ASR Systems
  868. Multilingual Speech Recognition with a Single End-to-End Model
  869. Multimodal Bag-of-Words for Cross Domains Sentiment Analysis
  870. Multimodal Signal Processing and Learning Aspects of Human-Robot Interaction for an Assistive Bathing Robot
  871. Multipitch Estimation Using Block Sparse Bayesian Learning and Intra-Block Clustering
  872. Multiple Feature Fusion for Automatic Emotion Recognition Using EEG Signals
  873. Multiple Jpeg Compression Detection Through Task-Driven Non-Negative Matrix Factorization
  874. Multiple Peer-to-Peer Bidirectional Cooperative Communications Using Massive MIMO Relays
  875. Multiple-Input Neural Network-Based Residual Echo Suppression
  876. Multiple-Model and Reduced-Order Kalman Filtering for Pathological Hand Tremor Extraction
  877. Multisource Mint Using Convolutive Transfer Function
  878. Multistream Diarization Fusion Using the Minimum Variance Bayesian Information Criterion
  879. Music Chord Recognition Based on Midi-Trained Deep Feature and BLSTM-CRF Hybird Decoding
  880. Music Structure Boundary Detection and Labelling by a Deconvolution of Path-Enhanced Self-Similarity Matrix
  881. Mutual-Information-Private Online Gradient Descent Algorithm
  882. Narrowband Channel Estimation for Hybrid Beamforming Millimeter Wave Communication Systems with One-Bit Quantization
  883. Nasal Speech Sounds Detection Using Connectionist Temporal Classification
  884. Natural TTS Synthesis by Conditioning Wavenet on MEL Spectrogram Predictions
  885. Nearest-Instance-Centroid-Estimation Linear Discriminant Analysis (Nice Lda)
  886. Negative Binomial Optimization for Biomedical Structural Variant Signal Reconstruction
  887. Neural Adaptive Image Denoiser
  888. Neural Confnet Classification: Fully Neural Network Based Spoken Utterance Classification Using Word Confusion Networks
  889. Neural Network Based Time-Frequency Masking and Steering Vector Estimation for Two-Channel Mvdr Beamforming
  890. Neural Network Language Modeling with Letter-Based Features and Importance Sampling
  891. Neural Sequential Malware Detection with Parameters
  892. New Multi-Carrier Demodulation Method Applied to Gearbox Vibration Analysis
  893. No Need for a Lexicon? Evaluating the Value of the Pronunciation Lexica in End-to-End Models
  894. No-Reference Hdr Image Quality Assessment Method Based on Tensor Space
  895. No-Reference Weighting Factor Selection for Bimodal Tomography
  896. Noise Robust Speech Recognition on Aurora4 by Humans and Machines
  897. Non-Asymptotic Guarantees for Correlation-Aware Support Detection
  898. Non-Euclidean Vector Product for Neural Networks
  899. Non-Iterative Missing Samples Recovery of ECG Signals by Lmmse Estimation for an Autoregressive Cyclostationary Model
  900. Non-Native Children Speech Recognition Through Transfer Learning
  901. Non-Negative Online Estimation for Hawkes Process Networks
  902. Non-Parallel Voice Conversion Using Variational Autoencoders Conditioned by Phonetic Posteriorgrams and D-Vectors
  903. Non-Zero Diffusion Particle Flow SMC-PHD Filter for Audio-Visual Multi-Speaker Tracking
  904. Noncircularity-Based Localization for Mixed Near-Field and Far-Field Sources with Unknown Mutual Coupling
  905. Nonconvex Sparse Logistic Regression via Proximal Gradient Descent
  906. Nonlinear Acoustic Echo Cancellation Using Elitist Resampling Particle Filter
  907. Nonlinear Speech Enhancement Under Speech PSD Uncertainty
  908. Nonnegative Matrix Factorization with Transform Learning
  909. Nonnegative Tensor Factorization for Source Separation of Loops in Audio
  910. Normalization of Partly Overlapping Audio Recordings from the Same Event Based on Relative Signal Powers
  911. Novel Algorithms for Exact and Efficient L1-NORM-BASED Tucker2 Decomposition
  912. Novel Bayesian Cluster Enumeration Criterion for Cluster Analysis with Finite Sample Penalty Term
  913. Novel Realizations of Speech-Driven Head Movements with Generative Adversarial Networks
  914. ON the Use of Wavenet as a Statistical Vocoder
  915. Object-Oriented Anomaly Detection in Surveillance Videos
  916. Oct Volumetric Data Restoration via Primal-Dual Plug-and-Play Method
  917. Octagonal-Axis Raster Pattern for Improved Test Zone Search Motion Estimation
  918. On Adversarial Training and Loss Functions for Speech Enhancement
  919. On Approximation of Bandlimited Functions with Compressed Sensing
  920. On Compressive Sensing of Sparse Covariance Matrices Using Deterministic Sensing Matrices
  921. On Consistency and Asymptotic Uniqueness in Quasi-Maximum Likelihood Blind Separation of Temporally-Diverse Sources
  922. On Error Resilient Design of Predictive Scalable Coding Systems
  923. On Information Coupling in Cooperative Network Synchronization
  924. On Maximum Likelihood Angle of Arrival Estimation Using Orthogonal Projections
  925. On Modular Training of Neural Acoustics-to-Word Model for LVCSR
  926. On SDW-MWF and Variable Span Linear Filter with Application to Speech Recognition in Noisy Environments
  927. On Selecting Antenna Placements in Indoor Radio Environments
  928. On Sequential Random Distortion Testing of Non-Stationary Processes
  929. On Spatial Features for Supervised Speech Separation and its Application to Beamforming and Robust ASR
  930. On Speech Enhancement Using Microphone Arrays in the Presence of Co-Directional Interference
  931. On Teaching Signals & Systems in a Project-Based Learning Environment
  932. On Using Backpropagation for Speech Texture Generation and Voice Conversion
  933. On the Analysis of Training Data for Wavenet-Based Speech Synthesis
  934. On the Comparison of Two Room Compensation / Dereverberation Methods Employing Active Acoustic Boundary Absorption
  935. On the Computability of System Approximations Under Causality Constraints
  936. On the Design of Robust Steerable Frequency-Invariant Beampatterns with Concentric Circular Microphone Arrays
  937. On the Equivalence of $f$-Divergence Balls and Density Bands in Robust Detection
  938. On the Geometry of Mixtures of Prescribed Distributions
  939. On the High-Snr Receiver Operating Characteristic of Glrt for The Conditional Signal Model
  940. On the Importance of Analytic Phase of Speech Signals in Spoken Language Recognition
  941. On the Modulus of Continuity for Noisy Positive Super-Resolution
  942. On the Performance Analysis of Wifi Based Localization
  943. On the Sample Complexity of Graphical Model Selection from Non-Stationary Samples
  944. On the Supermodularity of Active Graph-Based Semi-Supervised Learning with Stieltjes Matrix Regularization
  945. On the Use of Grapheme Models for Searching in Large Spoken Archives
  946. On-Talk and Off-Talk Detection: A Discrete Wavelet Transform Analysis of Electroencephalogram
  947. One-Bit Massive Mimo Precoding via a Minimum Symbol-Error Probability Design
  948. Online Direction of Arrival Estimation Based on Deep Learning
  949. Online Education Evaluation for Signal Processing Course Through Student Learning Pathways
  950. Online Multi-Kernel Learning with Orthogonal Random Features
  951. Open Set Recognition by Regularising Classifier with Fake Data Generated by Generative Adversarial Networks
  952. Opportunistic Sensing with MIC Arrays on Smart Speakers for Distal Interaction and Exercise Tracking
  953. Opportunistic Synchronisation of Multi-Static Staring Array Radars via Track-Before-Detect
  954. Optimal Algorithms and CRB for Reciprocity Calibration in Massive Mimo
  955. Optimal Crowdsourced Classification with a Reject Option in the Presence of Spammers
  956. Optimal Online Cyberbullying Detection
  957. Optimal Pooling of Covariance Matrix Estimates Across Multiple Classes
  958. Optimal Power Control Law for Equal-Rate DS-CDMA Networks Governed by a Successive Soft Interference Cancellation Scheme
  959. Optimal Power and Bit Allocation for Graph Signal Interpolation
  960. Optimal Selection of Subset of Images with Highest Intra-Class Similarity For 3D Scene Reconstruction
  961. Optimal Spectral Estimation and System Trade-Off in Long-Distance Frequency-Modulated Continuous-Wave Lidar
  962. Optimal Stopping Times for Estimating Bernoulli Parameters with Applications to Active Imaging
  963. Optimal Tone Reservation for Peak to Average Power Control of Cdma Systems
  964. Optimization of Speaker-Aware Multichannel Speech Extraction with ASR Criterion
  965. Optimized Sparse Array Design Based on the Sum Coarray
  966. Optimized Transmission for Consensus in Wireless Sensor Networks
  967. Optimizing Multilingual Knowledge Transfer for Time-Delay Neural Networks with Low-Rank Factorization
  968. Optimum Configurations of Sparse Subarray Beamformers
  969. Optimum Exact Histogram Specification
  970. Optimum Sparse Array Design for Maximizing Signal- to-Noise Ratio in Presence of Local Scatterings
  971. Optimum Sparse Array Design for Multiple Beamformers with Common Receiver
  972. Orthogonality-Regularized Masked NMF for Learning on Weakly Labeled Audio Data
  973. Orthogonally Regularized Deep Networks for Image Super-Resolution
  974. Out-of-Vocabulary Word Recovery using FST-Based Subword Unit Clustering in a Hybrid ASR System
  975. Outlier Removal for Enhancing Kernel-Based Classifier Via the Discriminant Information
  976. Overlapping Animal Sound Classification Using Sparse Representation
  977. PIPA: A New Proximal Interior Point Algorithm for Large-Scale Convex Optimization
  978. PVDC: A Binary Descriptor Using Pore-Valley Disk Code Structure for High-Resolution Partial Fingerprint Recognition
  979. Papr Minimization Through Spatio-Temporal Symbol-Level Precoding for the Non-Linear Multi-User MISO Channel
  980. Parallel Beamforming Design in Full Duplex Systems with Per-Antenna Power Constraints
  981. Parallel Stochastic Successive Convex Approximation Method for Large-Scale Dictionary Learning
  982. Parallel Vector Field Regularized Non-Negative Matrix Factorization for Image Representation
  983. Parallel-Data-Free Dictionary Learning for Voice Conversion Using Non-Negative Tucker Decomposition
  984. Parameter Estimation of Heavy-Tailed Random Walk Model from Incomplete Data
  985. Parameter Selection Strategy for Sparsity Enforcing Prior Models
  986. Parametric Approximation of Piano Sound Based on Kautz Model with Sparse Linear Prediction
  987. Particle Filtering and Inference for Limit Order Books in High Frequency Finance
  988. Particle Flow Particle Filter for Gaussian Mixture Noise Models
  989. Partitioning Relational Matrices of Similarities or Dissimilarities Using the Value of Information
  990. Pattern Localization in Time Series Through Signal-To-Model Alignment in Latent Space
  991. Perceptual Loss for Superpixel-Level Multispectral and Panchromatic Image Classification
  992. Perceptually Guided Speech Enhancement Using Deep Neural Networks
  993. Perceptually Motivated Analysis of Numerically Simulated Head-Related Transfer Functions Generated By Various 3D Surface Scanning Systems
  994. Performance of Interleaved Training for Single-User Hybrid Massive Antenna Downlink
  995. Performance of Mask Based Statistical Beamforming in a Smart Home Scenario
  996. Permissible Support Patterns for Identifying the Spreading Function of Time-Varying Channels
  997. Permutation Invariant Training for Speaker-Independent Multi-Pitch Tracking
  998. Permutation-Free Cgmm: Complex Gaussian Mixture Model with Inverse Wishart Mixture Model Based Spatial Prior for Permutation-Free Source Separation and Source Counting
  999. Phase Corrected Total Variation for Audio Signals
  1000. Phase Retrieval via Smoothing Projected Gradient Method

Looking for submission deadlines instead? See the conference deadline calendar.