← All conferences

ICASSP 2020 Accepted Papers

The full list of 1,849 papers accepted at ICASSP 2020 (IEEE International Conference on Acoustics, Speech and Signal Processing). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

  1. 1.5GBIT/S 4.9W Hyperspectral Image Encoders on a Low-Power Parallel Heterogeneous Processing Platform
  2. 2D-to-2D Mask Estimation for Speech Enhancement Based on Fully Convolutional Neural Network
  3. 3-D Acoustic Modeling for Far-Field Multi-Channel Speech Recognition
  4. 3D Unknown View Tomography Via Rotation Invariants
  5. 3d Deformation Signature for Dynamic Face Recognition
  6. A BI-Model Approach for Handling Unknown Slot Values in Dialogue State Tracking
  7. A Bidirectional Context Propagation Network for Urine Sediment Particle Detection in Microscopic Images
  8. A Bin Encoding Training of a Spiking Neural Network Based Voice Activity Detection
  9. A Comparative Study of Estimating Articulatory Movements from Phoneme Sequences and Acoustic Features
  10. A Comparative Study of Western and Chinese Classical Music Based on Soundscape Models
  11. A Comparison of Pooling Methods on LSTM Models for Rare Acoustic Event Classification
  12. A Complexity Efficient DMT-Optimal Tree Pruning Based Sphere Decoding
  13. A Composite DNN Architecture for Speech Enhancement
  14. A Comprehensive Framework for 2D-JND Extension to 360-DEG Images
  15. A Comprehensive Study of Residual CNNS for Acoustic Modeling in ASR
  16. A Computationally Light Algorithm for Bayesian Speech Enhancement with SNR Marginalization
  17. A Connected Auto-Encoders Based Approach for Image Separation with Side Information: With Applications to Art Investigation
  18. A Constrained Maximum Likelihood Estimator of Speech and Noise Spectra with Application to Multi-Microphone Noise Reduction
  19. A Cross-Task Transfer Learning Approach to Adapting Deep Speech Enhancement Models to Unseen Background Noise Using Paired Senone Classifiers
  20. A DSP Acceleration Framework For Software-Defined Radios On X86 64
  21. A Data Efficient End-to-End Spoken Language Understanding Architecture
  22. A Dataset for Measuring Reading Levels In India At Scale
  23. A Deep Gradient Boosting Network for Optic Disc and Cup Segmentation
  24. A Deep Learning Approach to Object Affordance Segmentation
  25. A Deep Learning Architecture for Epileptic Seizure Classification Based on Object and Action Recognition
  26. A Deep Multimodal Approach for Map Image Classification
  27. A Deep Neural Network-Driven Feature Learning Method for Polyphonic Acoustic Event Detection from Real-Life Recordings
  28. A Dense U-Net with Cross-Layer Intersection for Detection and Localization of Image Forgery
  29. A Dialogical Emotion Decoder for Speech Motion Recognition in Spoken Dialog
  30. A Differential Approach for Rain Field Tomographic Reconstruction Using Microwave Signals from Leo Satellites
  31. A Discriminative Condition-Aware Backend for Speaker Verification
  32. A Dual-Staged Context Aggregation Method towards Efficient End-to-End Speech Enhancement
  33. A Dynamic Stream Weight Backprop Kalman Filter for Audiovisual Speaker Tracking
  34. A Fast Non-Contact Vital Signs Detection Method Based on Regional Hidden Markov Model in A 77ghz Lfmcw Radar System
  35. A Fast Proximal Point Algorithm for Generalized Graph Laplacian Learning
  36. A Fast Reduced-Rank Sound Zone Control Algorithm Using The Conjugate Gradient Method
  37. A Fast Sparse Covariance-Based Fitting Method for DOA Estimation via Non-Negative Least Squares
  38. A Fast and Accurate Frequent Directions Algorithm for Low Rank Approximation via Block Krylov Iteration
  39. A Fast and Accurate Super-Resolution Network Using Progressive Residual Learning
  40. A Fifo Based Accelerator for Convolutional Neural Networks
  41. A Forward-Backward Algorithm for Reweighted Procedures: Application to Radio-Astronomical Imaging
  42. A Framework for Parameters Estimation of Image Operator Chain
  43. A Framework for the Robust Evaluation of Sound Event Detection
  44. A Frequency-Domain BSS Method Based on ℓ1 Norm, Unitary Constraint, and Cayley Transform
  45. A Gated Hypernet Decoder for Polar Codes
  46. A General Difficulty Control Algorithm for Proof-of-Work Based Blockchains
  47. A General Test for the Linear Structure of Covariance Matrices of Gaussian Populations
  48. A Generalization of Principal Component Analysis
  49. A Generalized Framework for Domain Adaptation of PLDA in Speaker Recognition
  50. A Geometric Approach for Unsupervised Similarity Learning
  51. A Graph Network Model for Distributed Learning with Limited Bandwidth Links and Privacy Constraints
  52. A Greedy Sparse Approximation Algorithm Based On L1-Norm Selection Rules
  53. A Hardware Architecture For Reconfigurable Intelligent Surfaces with Minimal Active Elements for Explicit Channel Estimation
  54. A Hierarchical Model for Dialog Act Recognition Considering Acoustic and Lexical Context Information
  55. A Hierarchical Tracker for Multi-Domain Dialogue State Tracking
  56. A Hybrid Approach for Thermographic Imaging With Deep Learning
  57. A Hybrid Model for Bipolar Disorder Classification from Visual Information
  58. A Hybrid Structural Sparse Error Model for Image Deblocking
  59. A Hybrid Text Normalization System Using Multi-Head Self-Attention For Mandarin
  60. A Large-Scale Deep Architecture for Personalized Grocery Basket Recommendations
  61. A Learning Approach to Cooperative Communication System Design
  62. A Lightweight Multi-Label Segmentation Network for Mobile Iris Biometrics
  63. A Linear Time Partitioning Algorithm for Frequency Weighted Impurity Functions
  64. A Low-Complexity Map Detector for Distributed Networks
  65. A Low-Dimensionality Method for Data-Driven Graph Learning
  66. A Low-Latency Successive Cancellation Hybrid Decoder for Convolutional Polar Codes
  67. A Low-Resolution ADC Proof-of-Concept Development for a Fully-Digital Millimeter-wave Joint Communication-Radar
  68. A Maximum Likelihood Approach to Multi-Objective Learning Using Generalized Gaussian Distributions for Dnn-Based Speech Enhancement
  69. A Memory Augmented Architecture for Continuous Speaker Identification in Meetings
  70. A Method for Millimeter-Wave Imaging of Concealed Objects Via De-Aliasing
  71. A Minimal Personalization of Dynamic Binaural Synthesis with Mixed Structural Modeling and Scattering Delay Networks
  72. A Model of Double Descent for High-Dimensional Logistic Regression
  73. A Model-Based Deep Network for MRI Reconstruction Using Approximate Message Passing Algorithm
  74. A Model-Free Approach to Distributed Transmit Beamforming
  75. A Moment-Based Approach for Guaranteed Tensor Decomposition
  76. A Monte Carlo Search-Based Triplet Sampling Method for Learning Disentangled Representation of Impulsive Noise on Steering Gear
  77. A Multi-Dilation and Multi-Resolution Fully Convolutional Network for Singing Melody Extraction
  78. A Multi-Phase Gammatone Filterbank for Speech Separation Via Tasnet
  79. A Multi-Scaled Receptive Field Learning Approach for Medical Image Segmentation
  80. A Multichannel Kalman-Based Wiener Filter Approach for Speaker Interference Reduction in Meetings
  81. A Multitaper Reassigned Spectrogram for Increased Time-Frequency Localization Precision
  82. A Neural Document Language Modeling Framework for Spoken Document Retrieval
  83. A Neural Network Based on First Principles
  84. A Neural Network for Monaural Intrusive Speech Intelligibility Prediction
  85. A Neural Network-Based Spike Sorting Feature Map That Resolves Spike Overlap in the Feature Space
  86. A New Application of Ultrasound Signal Processing for Archaeological Ceramic Classification
  87. A New Multihypothesis Prediction Scheme for Compressed Video Sensing Reconstruction
  88. A New Perspective for Flexible Feature Gathering in Scene Text Recognition Via Character Anchor Pooling
  89. A New Sampling Scheme for Distributed Blind Spectrum Sensing Using Energy Detectors
  90. A New Variational Method for Deep Supervised Semantic Image Hashing
  91. A Noninvasive Method to Detect Diabetes Mellitus and Lung Cancer Using the Stacked Sparse autoencoder
  92. A Novel Approach for Intelligibility Assessment in Dysarthric Subjects
  93. A Novel Method for Obtaining Diffuse Field Measurements for Microphone Calibration
  94. A Novel Moving Sparse Array Geometry with Increased Degrees of Freedom
  95. A Novel Pruning Approach for Bagging Ensemble Regression Based on Sparse Representation
  96. A Novel Rank Selection Scheme in Tensor Ring Decomposition Based on Reinforcement Learning for Deep Neural Networks
  97. A Novel Saliency-Driven Oil Tank Detection Method for Synthetic Aperture Radar Images
  98. A Novel Two-Pathway Encoder-Decoder Network for 3D Face Reconstruction
  99. A Partial Relaxation DOA Estimator Based on Orthogonal Matching Pursuit
  100. A Particle Gibbs Sampling Approach to Topology Inference in Gene Regulatory Networks
  101. A Practical Two-Stage Training Strategy for Multi-Stream End-to-End Speech Recognition
  102. A Priori Estimates of the Generalization Error for Autoencoders
  103. A Probabilistic Scheme for Representation Learning with Radial Transform Images
  104. A Prototypical Triplet Loss for Cover Detection
  105. A Proximal Dual Consensus Method for Linearly Coupled Multi-Agent Non-Convex Optimization
  106. A Random Gossip BMUF Process for Neural Language Modeling
  107. A Real Time Implementation of a Bayer Domain Image Deblurring Core for Optical Blur Compensation
  108. A Real-Time Deep Network for Crowd Counting
  109. A Recurrent Variational Autoencoder for Speech Enhancement
  110. A Recursive Bayesian Solution for the Excess Over Threshold Distribution with Stochastic Parameters
  111. A Recursive Edge Detector For Color Filter Array Image
  112. A Regularized Attention Mechanism for Graph Attention Networks
  113. A Return to Dereverberation in the Frequency Domain Using a Joint Learning Approach
  114. A Robust Audio-Visual Speech Enhancement Model
  115. A Robust Speaker Clustering Method Based on Discrete Tied Variational Autoencoder
  116. A Segmentation Based Robust Deep Learning Framework for Multimodal Retinal Image Registration
  117. A Self-Attentive Emotion Recognition Network
  118. A Semi-Supervised Approach For Identifying Abnormal Heart Sounds Using Variational Autoencoder
  119. A Semi-Supervised Rank Tracking Algorithm For On-Line Unmixing Of Hyperspectral Images
  120. A Sequence Matching Network for Polyphonic Sound Event Localization and Detection
  121. A Siamese Content-Attentive Graph Convolutional Network for Personality Recognition Using Physiology
  122. A Simple But Effective Bert Model for Dialog State Tracking on Resource-Limited Systems
  123. A Simple Derivation of AMP and its State Evolution via First-Order Cancellation
  124. A Simple and Efficient Iterative Method for Toa Localization
  125. A Single-RF Architecture for Multiuser Massive MIMO Via Reflecting Surfaces
  126. A Sparse Linear Array Approach in Automotive Radars Using Matrix Completion
  127. A Stacked-Autoencoder Based End-to-End Learning Framework for Decode-and-Forward Relay Networks
  128. A Streaming On-Device End-To-End Model Surpassing Server-Side Conventional Model Quality and Latency
  129. A Study of Child Speech Extraction Using Joint Speech Enhancement and Separation in Realistic Conditions
  130. A Study of Generalization of Stochastic Mirror Descent Algorithms on Overparameterized Nonlinear Models
  131. A Study on the Transferability of Adversarial Attacks in Sound Event Classification
  132. A Switching Transmission Game with Latency as the User's Communication Utility
  133. A Theoretical Basis for Practitioners Heuristic 1/N and Long-Only Quintile Portfolio
  134. A Time-Based Sampling Framework for Finite-Rate-of-Innovation Signals
  135. A Time-Frequency Network with Channel Attention and Non-Local Modules for Artificial Bandwidth Extension
  136. A Unified Sequence-to-Sequence Front-End Model for Mandarin Text-to-Speech Synthesis
  137. A Variational Bayesian Approach for Multichannel Through-Wall Radar Imaging with Low-Rank and Sparse Priors
  138. A Visual-Pilot Deep Fusion for Target Speech Separation in Multitalker Noisy Environment
  139. A Whiteness Test Based on the Spectral Measure of Large Non-Hermitian Random Matrices
  140. A WiFi-Based Passive Fall Detection System
  141. A Zeroth-Order Learning Algorithm for Ergodic Optimization of Wireless Systems with no Models and no Gradients
  142. A multi-view approach for Mandarin non-native mispronunciation verification
  143. A-CRNN: A Domain Adaptation Model for Sound Event Detection
  144. ADI17: A Fine-Grained Arabic Dialect Identification Dataset
  145. ADMM-Based One-Bit Quantized Signal Detection for Massive MIMO Systems With Hardware Impairments
  146. ADRN: Attention-Based Deep Residual Network for Hyperspectral Image Denoising
  147. AL2: Progressive Activation Loss for Learning General Representations in Classification Neural Networks
  148. APB2FACE: Audio-Guided Face Reenactment with Auxiliary Pose and Blink Signals
  149. ASR Error Correction and Domain Adaptation Using Machine Translation
  150. ASR is All You Need: Cross-Modal Distillation for Lip Reading
  151. AV(SE)2: Audio-Visual Squeeze-Excite Speech Enhancement
  152. Accelerating Distributed Deep Learning By Adaptive Gradient Quantization
  153. Accelerating Linear Algebra Kernels on a Massively Parallel Reconfigurable Architecture
  154. Accent Estimation of Japanese Words from Their Surfaces and Romanizations for Building Large Vocabulary Accent Dictionaries
  155. Accounting for Microprosody in Modeling Intonation
  156. Accuracy-Robustness Trade-Off for Positively Weighted Neural Networks
  157. Accurate 6D Object Pose Estimation by Pose Conditioned Mesh Reconstruction
  158. Accurate Localization of AUV in Motion by Explicit Solution Using Time Delays
  159. Accurate Semidefinite Relaxation Method for 3-D Rigid Body Localization Using AOA
  160. Accurate and Scalable Version Identification Using Musically-Motivated Embeddings
  161. Achieving Fully-Digital Performance by Hybrid Analog/Digital Beamforming in Wide-Band Massive-Mimo Systems
  162. Achieving the Capacity of the DNA Storage Channel
  163. Acoustic Matching By Embedding Impulse Responses
  164. Acoustic Model Adaptation for Presentation Transcription and Intelligent Meeting Assistant Systems
  165. Acoustic Scene Classification Using Deep Residual Networks with Late Fusion of Separated High and Low Frequency Paths
  166. Acoustic Scene Classification for Mismatched Recording Devices Using Heated-Up Softmax and Spectrum Correction
  167. Action-Manipulation Attacks on Stochastic Bandits
  168. Active Control of Line Spectral Noise with Simultaneous Secondary Path Modeling Without Auxiliary Noise
  169. Active Learning with Unsupervised Ensembles of Classifiers
  170. Active Noise Control Over Multiple Regions: Performance Analysis
  171. Active Semi-Supervised Learning for Diffusions on Graphs
  172. Acu-Net: A 3D Attention Context U-Net for Multiple Sclerosis Lesion Segmentation
  173. Adaptation and Learning in Multi-Task Decision Systems
  174. Adaptation of RNN Transducer with Text-To-Speech Technology for Keyword Spotting
  175. Adaptive Blind Audio Source Extraction Supervised By Dominant Speaker Identification Using X-Vectors
  176. Adaptive Distributed Stochastic Gradient Descent for Minimizing Delay in the Presence of Stragglers
  177. Adaptive Elastic Loss Based on Progressive Inter-Class Association for Cervical Histology Image Segmentation
  178. Adaptive Knowledge Distillation Based on Entropy
  179. Adaptive Matched Filter using Non-Target Free Training Data
  180. Adaptive Normalization for Forecasting Limit Order Book Data Using Convolutional Neural Networks
  181. Adaptive Prediction of Financial Time-Series for Decision-Making Using A Tensorial Aggregation Approach
  182. Adaptive Region Aggregation Network: Unsupervised Domain Adaptation with Adversarial Training for ECG Delineation
  183. Adaptive Sequential Interpolator Using Active Learning for Efficient Emulation of Complex Systems
  184. Adaptive Subspace Detectors for off-grid Mismatched Targets
  185. Addressing Accent Mismatch In Mandarin-English Code-Switching Speech Recognition
  186. Addressing Challenges in Building Web-Scale Content Classification Systems
  187. Addressing The Confounds Of Accompaniments In Singer Identification
  188. Addressing the Polysemy Problem in Language Modeling with Attentional Multi-Sense Embeddings
  189. AdvMS: A Multi-Source Multi-Cost Defense Against Adversarial Attacks
  190. Adversarial Anomaly Detection for Marked Spatio-Temporal Streaming Data
  191. Adversarial Attacks on Deep Unfolded Networks for Sparse Coding
  192. Adversarial Attacks on GMM I-Vector Based Speaker Verification Systems
  193. Adversarial Detection of Counterfeited Printable Graphical Codes: Towards "Adversarial Games" In Physical World
  194. Adversarial Example Detection by Classification for Deep Speech Recognition
  195. Adversarial Mixup Synthesis Training for Unsupervised Domain Adaptation
  196. Adversarial Multi-Task Learning for Speaker Normalization in Replay Detection
  197. Adversarial Networks for Secure Wireless Communications
  198. Adversarial Text Image Super-Resolution using Sinkhorn Distance
  199. Adversarial Video Compression Guided by Soft Edge Detection
  200. Age of Information with Finite Horizon and Partial Updates
  201. Age-Based Scheduling Policy for Federated Learning in Mobile Edge Networks
  202. Aipnet: Generative Adversarial Pre-Training of Accent-Invariant Networks for End-To-End Speech Recognition
  203. Algorithmic Exploration of American English Dialects
  204. Alignment-Length Synchronous Decoding for RNN Transducer
  205. Aligntts: Efficient Feed-Forward Text-to-Speech System Without Explicit Alignment
  206. All In One Network for Driver Attention Monitoring
  207. All You Need is a Second Look: Towards Tighter Arbitrary Shape Text Detection
  208. Allocation of Computing Tasks In Distributed MEC Servers Co-Powered By Renewable Sources And The Power Grid
  209. Alternative Half-Sample Interpolation Filters for Versatile Video Coding
  210. An Acoustic Modelling Based Remote Error Sensing Approach for Quiet Zone Generation in a Noisy Environment
  211. An Adaptive Linear Estimator Based Approach to Bi-Directional Motion Compensated Prediction
  212. An Alternative Signature Design Using L1 Principal Components for Spread-Spectrum Steganography
  213. An Analysis of Speech Enhancement and Recognition Losses in Limited Resources Multi-Talker Single Channel Audio-Visual ASR
  214. An Analytical Solution to Jacobsen Estimator for Windowed Signals
  215. An Attention Enhanced Multi-Task Model for Objective Speech Assessment in Real-World Environments
  216. An Attention-Based Joint Acoustic and Text on-Device End-To-End Model
  217. An Early Termination Scheme for Successive Cancellation List Decoding of Polar Codes
  218. An Easy-to-Implement Framework of Fast Subspace Clustering For Big Data Sets
  219. An Efficient Augmented Lagrangian-Based Method for Linear Equality-Constrained Lasso
  220. An Efficient Methodology to De-Anonymize the 5G-New Radio Physical Downlink Control Channel
  221. An Empirical Bayes Approach to Partially Labeled and Shuffled Data Sets
  222. An Empirical Study of Conv-Tasnet
  223. An Empirical Study of Transformer-Based Neural Language Model Adaptation
  224. An Empirical Study on Acoustic Feedback Path Across Hearing Aid Users
  225. An Enhanced Decoding Algorithm for Coded Compressed Sensing
  226. An Ensemble Based Approach for Generalized Detection of Spoofing Attacks to Automatic Speaker Recognizers
  227. An Improved Deep Neural Network for Modeling Speaker Characteristics at Different Temporal Scales
  228. An Improved Frame-Unit-Selection Based Voice Conversion System Without Parallel Training Data
  229. An Improved Selective Active Noise Control Algorithm Based on Empirical Wavelet Transform
  230. An Improved Solution to the Frequency-Invariant Beamforming with Concentric Circular Microphone Arrays
  231. An LSTM Based Architecture to Relate Speech Stimulus to Eeg
  232. An LSTM-Based Dynamic Chord Progression Generation System for Interactive Music Performance
  233. An Odorant Encoding Machine for Sampling, Reconstruction and Robust Representation of Odorant Identity
  234. An Online Kernel Scalar Quantization Scheme for Signal Classification
  235. An Online Speaker-aware Speech Separation Approach Based on Time-domain Representation
  236. An Ontology-Aware Framework for Audio Event Classification
  237. An Optimal Channel Estimation Scheme for Intelligent Reflecting Surfaces Based on a Minimum Variance Unbiased Estimator
  238. An Optimal Symmetric Threshold Strategy for Remote Estimation Over The Collision Channel
  239. An Unsupervised Retinal Vessel Extraction and Segmentation Method Based On a Tube Marked Point Process Model
  240. Analysis of Acoustic Features for Speech Sound Based Classification of Asthmatic and Healthy Subjects
  241. Analyzing ASR Pretraining for Low-Resource Speech-to-Text Translation
  242. Anefficient Alternative to Network Pruning Through Ensemble Learning
  243. Angular Discriminative Deep Feature Learning for Face Verification
  244. Anomalous Sound Detection Based on Interpolation Deep Neural Network
  245. Anomaly Detection for Time Series Using VAE-LSTM Hybrid Model
  246. Anomaly Detection in Mixed Time-Series Using A Convolutional Sparse Representation With Application To Spacecraft Health Monitoring
  247. Anomaly Detection with Training Data in Hyperspectral Imagery
  248. Anomalydae: Dual Autoencoder for Anomaly Detection on Attributed Networks
  249. Anti-Jamming Routing For Internet of Satellites: a Reinforcement Learning Approach
  250. Anytime Minibatch with Delayed Gradients: System Performance and Convergence Analysis
  251. Application Informed Motion Signal Processing for Finger Motion Tracking Using Wearable Sensors
  252. Approaching Optimal Embedding In Audio Steganography With GAN
  253. Approximate Bayesian Computation with the Sliced-Wasserstein Distance
  254. Approximate Inference by Kullback-Leibler Tensor Belief Propagation
  255. Arnet: Attention-Based Refinement Network for Few-Shot Semantic Segmentation
  256. Array-Geometry-Aware Spatial Active Noise Control Based on Direction-of-Arrival Weighting
  257. Arsm Gradient Estimator for Supervised Learning to Rank
  258. Artificial Bandwidth Extension Using Conditional Variational Auto-encoders and Adversarial Learning
  259. Assessing the Scope of Generalized Countermeasures for Anti-Spoofing
  260. Assimilation-Based Learning of Chaotic Dynamical Systems from Noisy and Partial Data
  261. Asymptotic Stochastic Analysis of Partially Relaxed DML
  262. Asymptotically Optimal Blind Calibration of Acoustic Vector Sensor Uniform Linear Arrays
  263. Asynchrounous Decentralized Learning of a Neural Network
  264. Atomic Norm Based Localization of Far-Field and Near-Field Signals with Generalized Symmetric Arrays
  265. Atomic Norm Denoising In Blind Two-Dimensional Super-Resolution
  266. Atrial Fibrillation Risk Prediction from Electrocardiogram and Related Health Data with Deep Neural Network
  267. Attention Driven Fusion for Multi-Modal Emotion Recognition
  268. Attention Guided Region Division for Crowd Counting
  269. Attention Mechanism Enhanced Kernel Prediction Networks for Denoising of Burst Images
  270. Attention-Based ASR with Lightweight and Dynamic Convolutions
  271. Attention-Based Curiosity-Driven Exploration in Deep Reinforcement Learning
  272. Attention-Based Gated Scaling Adaptive Acoustic Model for CTC-Based Speech Recognition
  273. Attention-Guided Deraining Network Via Stage-Wise Learning
  274. Attention-Mask Dense Merger (Attendense) Deep HDR for Ghost Removal
  275. Attentional Fused Temporal Transformation Network for Video Action Recognition
  276. Attentive Cutmix: An Enhanced Data Augmentation Approach for Deep Learning Based Image Classification
  277. Attentive Item2vec: Neural Attentive User Representations
  278. Attentive Modality Hopping Mechanism for Speech Emotion Recognition
  279. Audio Codec Enhancement with Generative Adversarial Networks
  280. Audio Feature Extraction for Vehicle Engine Noise Classification
  281. Audio Sound Determination Using Feature Space Attention Based Convolution Recurrent Neural Network
  282. Audio-Assisted Image Inpainting for Talking Faces
  283. Audio-Attention Discriminative Language Model for ASR Rescoring
  284. Audio-Based Auto-Tagging With Contextual Tags for Music
  285. Audio-Based Detection of Explicit Content in Music
  286. Audio-Visual Calibration with Polynomial Regression for 2-D Projection Using SVD-PHAT
  287. Audio-Visual Recognition of Overlapped Speech for the LRS2 Dataset
  288. Auditory Model Based Subsetting of Head-Related Transfer Function Datasets
  289. Auglabel: Exploiting Word Representations to Augment Labels for Face Attribute Classification
  290. Augmentation Data Synthesis Via Gans: Boosting Latent Fingerprint Reconstruction
  291. Augmented Grad-CAM: Heat-Maps Super Resolution Through Augmentation
  292. Augmenting Molecular Images with Vector Representations as a Featurization Technique for Drug Classification
  293. Auto-Fas: Searching Lightweight Networks for Face Anti-Spoofing
  294. Automatic Classification of Volumes of Water Using Swallow Sounds from Cervical Auscultation
  295. Automatic Data Augmentation Via Deep Reinforcement Learning for Effective Kidney Tumor Segmentation
  296. Automatic Epileptic Seizure Onset-Offset Detection Based On CNN in Scalp EEG
  297. Automatic Event Detection of REM Sleep Without Atonia From Polysomnography Signals Using Deep Neural Networks
  298. Automatic Fluency Evaluation of Spontaneous Speech Using Disfluency-Based Features
  299. Automatic Identification of Speakers From Head Gestures in a Narration
  300. Automatic Lyrics Alignment and Transcription in Polyphonic Music: Does Background Music Help?
  301. Automatic Prediction of Suicidal Risk in Military Couples Using Multimodal Interaction Cues from Couples Conversations
  302. Automatic and Simultaneous Adjustment of Learning Rate and Momentum for Stochastic Gradient-based Optimization Methods
  303. Automotive Collision Risk Estimation Under Cooperative Sensing
  304. Automotive Radar Signal Interference Mitigation Using RNN with Self Attention
  305. Autoregressive Parameter Estimation with Dnn-Based Pre-Processing
  306. Auxiliary Capsules for Natural Language Understanding
  307. Ava Active Speaker: An Audio-Visual Dataset for Active Speaker Detection
  308. BBA-NET: A Bi-Branch Attention Network For Crowd Counting
  309. BBAND INDEX: A NO-REFERENCE BANDING ARTIFACT PREDICTOR
  310. BOFFIN TTS: Few-Shot Speaker Adaptation by Bayesian Optimization
  311. BP-VB-EP Based Static and Dynamic Sparse Bayesian Learning with Kronecker Structured Dictionaries
  312. Back-And-Forth Prediction for Deep Tensor Compression
  313. Back-to-Back Butterfly Network, an Adaptive Permutation Network for New Communication Standards
  314. Balanced Binary Neural Networks with Gated Residual
  315. Balancing Rates and Variance via Adaptive Batch-Sizes in First-Order Stochastic Optimization
  316. Bandit Sampling for Faster Activity and Data Detection in Massive Random Access
  317. Bandwidth Extension of Musical Audio Signals With No Side Information Using Dilated Convolutional Neural Networks
  318. Bangla Voice Command Recognition in end-to-end System Using Topic Modeling based Contextual Rescoring
  319. Batman: Bayesian Target Modelling For Active Inference
  320. Bayesian Estimation of Plda with Noisy Training Labels, with Applications to Speaker Verification
  321. Bayesian Multiple Change-Point Detection with Limited Communication
  322. Beam Elimination Based on Sequentially Estimated a Posteriori Probabilities of Winning
  323. Beam-TasNet: Time-domain Audio Separation Network Meets Frequency-domain Beamformer
  324. Beamformed Feature for Learning-based Dual-channel Speech Separation
  325. Beamforming Design for High-Resolution Low-Intensity Focused Ultrasound Neuromodulation
  326. Beamforming in Intelligent Environments based on Ultra-Massive MIMO Platforms in Millimeter Wave and Terahertz Bands
  327. Bert is Not All You Need for Commonsense Inference
  328. Better Safe Than Sorry: Risk-Aware Nonlinear Bayesian Estimation
  329. Beyond the Dcase 2017 Challenge on Rare Sound Event Detection: A Proposal for a More Realistic Training and Test Framework
  330. Bilateral Recurrent Network for Single Image Deraining
  331. Binary Probability Model for Learning Based Image Compression
  332. Binaural Audio Source Remixing with Microphone Array Listening Devices
  333. Bio-Mimetic Attentional Feedback in Music Source Separation
  334. Bipartite Belief Propagation Polar Decoding With Bit-Flipping
  335. Bit Allocation for Multi-Task Collaborative Intelligence
  336. Blaster: An Off-Grid Method for Blind and Regularized Acoustic Echoes Retrieval
  337. Blind Bounded Source Separation Using Neural Networks with Local Learning Rules
  338. Blind Hyperspectral Unmixing using Dual Branch Deep Autoencoder with Orthogonal Sparse Prior
  339. Blind Inference of Centrality Rankings from Graph Signals
  340. Blind Multi-Spectral Image Pan-Sharpening
  341. Blind Source Separation of Graph Signals
  342. Blood Pressure Estimation From PPG Signals Using Convolutional Neural Networks And Siamese Network
  343. Body Movement Generation for Expressive Violin Performance Applying Neural Networks
  344. Boosted Locality Sensitive Hashing: Discriminative Binary Codes for Source Separation
  345. Breathing and Speech Planning in Spontaneous Speech Synthesis
  346. Bridging Mixture Density Networks with Meta-Learning for Automatic Speaker Identification
  347. Bringing in the Outliers: A Sparse Subspace Clustering Approach to Learn a Dictionary of Mouse Ultrasonic Vocalizations
  348. Building Firmly Nonexpansive Convolutional Neural Networks
  349. But System for the Second Dihard Speech Diarization Challenge
  350. Byzantine-Robust Decentralized Stochastic Optimization
  351. C3DVQA: Full-Reference Video Quality Assessment with 3D Convolutional Neural Network
  352. CAD-AEC: Context-Aware Deep Acoustic Echo Cancellation
  353. CGCNN: Complex Gabor Convolutional Neural Network on Raw Speech
  354. CIF: Continuous Integrate-And-Fire for End-To-End Speech Recognition
  355. CLCNET: Deep Learning-Based Noise Reduction for Hearing aids using Complex Linear Coding
  356. CN-Celeb: A Challenging Chinese Speaker Recognition Dataset
  357. CNN-Based Analog CSI Feedback in FDD MIMO-OFDM Systems
  358. CORRGAN: Sampling Realistic Financial Correlation Matrices Using Generative Adversarial Networks
  359. CP-GAN: Context Pyramid Generative Adversarial Network for Speech Enhancement
  360. CPWC: Contextual Point Wise Convolution for Object Recognition
  361. CS-R-FCN: Cross-Supervised Learning for Large-Scale Object Detection
  362. Camera Configuration Design in Cooperative Active Visual 3d Reconstruction: A Statistical Approach
  363. Can every analog system be simulated on a digital computer?
  364. Capacity of the Erasure Shuffling Channel
  365. Cartoon-Texture Decomposition-Based Variational Pansharpening
  366. Cell-Phone Classification: A Convolutional Neural Network Approach Exploiting Electromagnetic Emanations
  367. Challenges and Perspectives in Neuromorphic-based Visual IoT Systems and Networks
  368. Channel Adversarial Training for Speaker Verification and Diarization
  369. Channel Attention Based Generative Network for Robust Visual Tracking
  370. Channel Charting: an Euclidean Distance Matrix Completion Perspective
  371. Channel Covariance Estimation in Multiuser Massive Mimo Systems with an Approach Based on Infinite Dimensional Hilbert Spaces
  372. Channel Invariant Speaker Embedding Learning with Joint Multi-Task and Adversarial Training
  373. Channel Selection over Riemannian Manifold with Non-Stationarity Consideration for Brain-Computer Interface Applications
  374. Channel-Attention Dense U-Net for Multichannel Speech Enhancement
  375. Characterisation of a Snapshot Fourier Transform Imaging Spectrometer Based on an Array of Fabry-Perot Interferometers
  376. Characterizing Speech Adversarial Examples Using Self-Attention U-Net Enhancement
  377. Chirping up the Right Tree: Incorporating Biological Taxonomies into Deep Bioacoustic Classifiers
  378. Classification of Depth and Surface Edges with Deep Features
  379. Classification of Epileptic IEEG Signals by CNN and Data Augmentation
  380. Classification of High-Dimensional Motor Imagery Tasks Based on An End-To-End Role Assigned Convolutional Neural Network
  381. Classify and Explain: An Interpretable Convolutional Neural Network For Lung Cancer Diagnosis
  382. Classifying Anomalies for Network Security
  383. Classifying Partially Labeled Networked Data VIA Logistic Network Lasso
  384. Clock Synchronization Over Networks Using Sawtooth Models
  385. Clotho: an Audio Captioning Dataset
  386. Cloud-Driven Multi-Way Multiple-Antenna Relay Systems: Best-User-Link Selection and Joint Mmse Detection
  387. Clustering of Nonnegative Data and an Application to Matrix Completion
  388. Clutter Identification Based on Sparse Recovery and L1-Type Probabilistic Distance Measures
  389. Cochlear Signal Processing: A Platform for Learning the Fundamentals of Digital Signal Processing
  390. Code-Switched Speech Synthesis Using Bilingual Phonetic Posteriorgram with Only Monolingual Corpora
  391. Coded Illumination and Multiplexing for Lensless Imaging
  392. Cogans For Unsupervised Visual Speech Adaptation To New Speakers
  393. Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision
  394. Color Stabilization for Multi-Camera Light-Field Imaging
  395. Color and Angular Reconstruction of Light Fields from Incomplete-Color Coded Projections
  396. Colour Compression of Plenoptic Point Clouds Using Raht-Klt with Prior Colour Clustering and Specular/Diffuse Component Separation
  397. Combining Acoustics, Content and Interaction Features to Find Hot Spots in Meetings
  398. Combining CGAN and Mil for Hotspot Segmentation in Bone Scintigraphy
  399. Combining Deep Embeddings of Acoustic and Articulatory Features for Speaker Identification
  400. Communication Constrained Learning with Uncertain Models
  401. Commuting Conditional GANS for Multi-Modal Fusion
  402. Compare Learning: Bi-Attention Network for Few-Shot Learning
  403. Comparison of Glottal Closure Instants Detection Algorithms for Emotional Speech
  404. Comparison of User Models Based on GMM-UBM and I-Vectors for Speech, Handwriting, and Gait Assessment of Parkinson's Disease Patients
  405. Complex Pairwise Activity Analysis Via Instance Level Evolution Reasoning
  406. Complex Trainable Ista for Linear and Nonlinear Inverse Problems
  407. Complex Transformer: A Framework for Modeling Complex-Valued Sequence
  408. Complexity Reduction Methods for Index Modulation Based Dual-Function Radar Communication Systems
  409. Composite Dynamic Texture Synthesis Using Hierarchical Linear Dynamical System
  410. Compressed Sensing Based Channel Estimation and Open-loop Training Design for Hybrid Analog-digital Massive MIMO Systems
  411. Compressing Flow Fields with Edge-Aware Homogeneous Diffusion Inpainting
  412. Compressive 2-d Off-grid DOA Estimation for Propeller Cavitation Localization
  413. Compressive Adaptive Bilateral Filtering
  414. Computability of the Peak Value of Bandlimited Signals
  415. Computation of "Best" Interpolants in the Lp Sense
  416. Computing Hilbert Transform and Spectral Factorization for Signal Spaces of Smooth Functions
  417. Concentration-Based Polynomial Calculations on Nicked DNA
  418. Conditional Density Driven Grid Design in Point-Mass Filter
  419. Conditional Domain Adversarial Transfer for Robust Cross-Site ADHD Classification Using Functional MRI
  420. Conditional Mutual Information Neural Estimator
  421. Confidence Estimation for Black Box Automatic Speech Recognition Systems Using Lattice Recurrent Neural Networks
  422. Confirmnet: Convolutional Firmnet and Application to Image Denoising and Inpainting
  423. Consensus-Based Distributed Clustering for IoT
  424. Consistency-Aware Multi-Channel Speech Enhancement Using Deep Neural Networks
  425. Constant Envelope Massive MIMO-OFDM Precoding: an Improved Formulation and Solution
  426. Constant-Envelope Precoding for Satellite Systems
  427. Constrained Spectral Clustering for Dynamic Community Detection
  428. Content Based Singing Voice Extraction from a Musical Mixture
  429. Content Vs Context: How About "Walking Hand-In-Hand" For Image Clustering?
  430. Context and Uncertainty Modeling for Online Speaker Change Detection
  431. Continual Learning Through One-Class Classification Using VAE
  432. Continual Learning for Infinite Hierarchical Change-Point Detection
  433. Continuous Speech Separation: Dataset and Analysis
  434. Control of Linear Dynamical Systems Using Sparse Inputs
  435. Controllable Time-Delay Transformer for Real-Time Punctuation Prediction and Disfluency Detection
  436. Controlling the Perceived Sound Quality for Dialogue Enhancement With Deep Learning
  437. Convergence-Guaranteed Independent Positive Semidefinite Tensor Analysis Based on Student's T Distribution
  438. Converting Written Language to Spoken Language with Neural Machine Translation for Language Modeling
  439. Convex Optimisation-Based Privacy-Preserving Distributed Average Consensus in Wireless Sensor Networks
  440. Convolutional Beamspace for Array Signal Processing
  441. Cooperative Learning VIA Federated Distillation OVER Fading Channels
  442. Corrdrop: Correlation Based Dropout for Convolutional Neural Networks
  443. Correction of Automatic Speech Recognition with Transformer Sequence-To-Sequence Model
  444. Correlated Multi-Armed Bandits with A Latent Random Source
  445. Cost Aware Adversarial Learning
  446. Counting Dense Objects in Remote Sensing Images
  447. Coupled Training of Sequence-to-Sequence Models for Accented Speech Recognition
  448. Cra: A Generic Compression Ratio Adapter for End-To-End Data-Driven Image Compressive Sensing Reconstruction Frameworks
  449. Cramer-Rao Bound on DOA Estimation of Finite Bandwidth Signals Using a Moving Sensor
  450. Cramér-Rao Bounds for Flaw Localization in Subsampled Multistatic Multichannel Ultrasound Ndt Data
  451. Crnn-Ctc Based Mandarin Keywords Spotting
  452. Cross Image Cubic Interpolator for Spatially Varying Exposures
  453. Cross Lingual Transfer Learning for Zero-Resource Domain Adaptation
  454. Cross-Domain Adaptation for Biometric Identification Using Photoplethysmogram
  455. Cross-Domain Joint Dictionary Learning for ECG Reconstruction from PPG
  456. Cross-Lingual Topic Prediction For Speech Using Translations
  457. Cross-Speaker Silent-Speech Command Word Recognition Using Electro-Optical Stomatography
  458. Cross-Stained Segmentation from Renal Biopsy Images Using Multi-Level Adversarial Learning
  459. Cross-VAE: Towards Disentangling Expression from Identity For Human Faces
  460. Cross-View Attention Network for Breast Cancer Screening from Multi-View Mammograms
  461. Crowdsourcing-Based Ranking Aggregation for Person Re-Identification
  462. Cumulant Slice Reconstruction from Compressive Measurements and Its Application to Line Spectrum Estimation
  463. D-SLAM: Diffusion Source Localization and Trajectory Mapping
  464. D2NA: Day-To-Night Adaptation for Vision based Parking Management System
  465. DEJA-VU: Double Feature Presentation and Iterated Loss in Deep Transformer Networks
  466. DGAN: Disentangled Representation Learning for Anisotropic BRDF Reconstruction
  467. DNN-Based Speech Presence Probability Estimation for Multi-Frame Single-Microphone Speech Enhancement
  468. DNN-Based Speech Recognition for Globalphone Languages
  469. DNN-Chip Predictor: An Analytical Performance Predictor for DNN Accelerators with Various Dataflows and Hardware Architectures
  470. DNN-based Distributed Multichannel Mask Estimation for Speech Enhancement in Microphone Arrays
  471. DNN-based Mask Estimation Integrating Spectral and Spatial Features for Robust Beamforming
  472. DNN-supported Mask-based Convolutional Beamforming for Simultaneous Denoising, Dereverberation, and Source Separation
  473. DOA Estimation in Systems with Nonlinearities for MMWAVE Communications
  474. DOA Tracking Via Signal-Subspace Projector Update
  475. Damage-Sensitive and Domain-Invariant Feature Extraction for Vehicle-Vibration-Based Bridge Health Monitoring
  476. Data Augmentation Using Empirical Mode Decomposition on Neural Networks to Classify Impact Noise in Vehicle
  477. Data Selection Kernel Conjugate Gradient Algorithm
  478. Data-Driven Harmonic Filters for Audio Representation Learning
  479. Data-Driven Model Set Design for Model Averaged Particle Filter
  480. Data-Driven Wind Speed Estimation Using Multiple Microphones
  481. Deblurring And Super-Resolution Using Deep Gated Fusion Attention Networks For Face Images
  482. Decentralized Min-Max Optimization: Formulations, Algorithms and Applications in Network Poisoning Attack
  483. Decentralized Optimization with Non-Identical Sampling in Presence of Stragglers
  484. Decentralized Stochastic Non-Convex Optimization over Weakly Connected Time-Varying Digraphs
  485. Decentralized expected consistent signal recovery for quantization Measurements
  486. Decidable Variable-Rate Dataflow for Heterogeneous Signal Processing Systems
  487. Decoding 5G-NR Communications VIA Deep Learning
  488. Decoding Movement Imagination and Execution From Eeg Signals Using Bci-Transfer Learning Method Based on Relation Network
  489. Decomposed Cyclegan for Single Image Deraining With Unpaired Data
  490. Deep Audio-Visual Speech Separation with Attention Mechanism
  491. Deep Autotuner: A Pitch Correcting Network for Singing Performances
  492. Deep Casa for Talker-independent Monaural Speech Separation
  493. Deep Clustering for Domain Adaptation
  494. Deep Clusteringwith Concrete K-Means
  495. Deep Contextualized Acoustic Representations for Semi-Supervised Speech Recognition
  496. Deep Encoded Linguistic and Acoustic Cues for Attention Based End to End Speech Emotion Recognition
  497. Deep Exposure Fusion with Deghosting via Homography Estimation and Attention Learning
  498. Deep Flow Collaborative Network for Online Visual Tracking
  499. Deep Geometric Knowledge Distillation with Graphs
  500. Deep Image Deblurring Using Local Correlation Block
  501. Deep James-Stein Neural Networks For Brain-Computer Interfaces
  502. Deep Joint Source-Channel Coding for Wireless Image Retrieval
  503. Deep Joint Source-Channel Coding of Images with Feedback
  504. Deep Learning Abilities to Classify Intricate Variations in Temporal Dynamics of Multivariate Time Series
  505. Deep Learning Based Prediction of Hypernasality for Clinical Applications
  506. Deep Learning for Robust Power Control for Wireless Networks
  507. Deep Learning-Based Beam Alignment in Mmwave Vehicular Networks
  508. Deep Matrix Completion on Graphs: Application in Drug Target Interaction Prediction
  509. Deep Meta-Relation Network for Visual Few-Shot Learning
  510. Deep Metric Learning Based On Center-Ranked Loss for Gait Recognition
  511. Deep Monocular Video Depth Estimation Using Temporal Attention
  512. Deep Multi-Region Hashing
  513. Deep Multi-Scale Gabor Wavelet Network for Image Restoration
  514. Deep Neural Network Based Matrix Completion for Internet of Things Network Localization
  515. Deep Neural Networks Based Automatic Speech Recognition for Four Ethiopian Languages
  516. Deep Product Quantization Module for Efficient Image Retrieval
  517. Deep Rainrate Estimation from Highly Attenuated Downlink Signals of Ground-Based Communications Satellite Terminals
  518. Deep Residual Network for MSFA Raw Image Denoising
  519. Deep Soft Interference Cancellation for MIMO Detection
  520. Deep Speech Extraction with Time-Varying Spatial Filtering Guided By Desired Direction Attractor
  521. Deep-Neural-Network Based Fall-Back Mechanism in Interference-Aware Receiver Design
  522. Deep-SST-Eddies: A Deep Learning Framework to Detect Oceanic Eddies in Sea Surface Temperature Images
  523. Defending Graph Convolutional Networks Against Adversarial Attacks
  524. Defense Against Adversarial Attacks on Spoofing Countermeasures of ASV
  525. Deliberation Model Based Two-Pass End-To-End Speech Recognition
  526. Demystifying TasNet: A Dissecting Approach
  527. Denoising of Event-Based Sensors with Spatial-Temporal Correlation
  528. Dense Mapping of Intracellular Diffusion and Drift from Single-Particle Tracking Data Analysis
  529. Dense Residual Network for Retinal Vessel Segmentation
  530. Densely Connected Neural Network with Dilated Convolutions for Real-Time Speech Enhancement in The Time Domain
  531. Depth Estimation From Single Image Through Multi-Path-Multi-Rate Diverse Feature Extractor
  532. Depth Map Fingerprinting and Splicing Detection
  533. Depthwise-STFT Based Separable Convolutional Neural Networks
  534. Deriving Compact Feature Representations Via Annealed Contraction
  535. Design Considerations for Hypothesis Rejection Modules in Spoken Language Understanding Systems
  536. Design of A Convergence-Aware Based Expectation Propagation Algorithm for Uplink Mimo Scma Systems
  537. Design-Gan: Cross-Category Fashion Translation Driven By Landmark Attention
  538. Detect Insider Attacks Using CNN in Decentralized Optimization
  539. Detecting Adversarial Attacks In Time-Series Data
  540. Detecting Autism Spectrum Disorder Using Topological Data Analysis
  541. Detecting Emotion Primitives from Speech and Their Use in Discerning Categorical Emotions
  542. Detecting Mismatch Between Text Script and Voice-Over Using Utterance Verification Based on Phoneme Recognition Ranking
  543. Detecting Multiple Speech Disfluencies Using a Deep Residual Network with Bidirectional Long Short-Term Memory
  544. Detection Of S1 And S2 Locations In Phonocardiogram Signals Using Zero Frequency Filter
  545. Detection and Analysis of T/D Deletion in Librispeech
  546. Detection of Adversarial Attacks and Characterization of Adversarial Subspace
  547. Detection of Malicious Vbscript Using Static and Dynamic Analysis with Recurrent Deep Learning
  548. Detection of Mild Dyspnea from Pairs of Speech Recordings
  549. Detection of Speech Events and Speaker Characteristics through Photo-Plethysmographic Signal Neural Processing
  550. Determined Source Separation Using the Sparsity of Impulse Responses
  551. Deterministic Feature Decoupling by Surfing Invariance Manifolds
  552. Dfsmn-San with Persistent Memory Model for Automatic Speech Recognition
  553. Diacritic-Level Pronunciation Analysis Using Phonological Features
  554. Diagonalizable Shift and Filters for Directed Graphs Based on the Jordan-Chevalley Decomposition
  555. Dialogue History Integration into End-to-End Signal-to-Concept Spoken Language Understanding Systems
  556. Differentiable Branching In Deep Networks for Fast Inference
  557. Digital Watermarking For Protecting Audio Classification Datasets
  558. Dilated Convolutional Neural Networks for Panoramic Image Saliency Prediction
  559. Discovering Causalities from Cardiotocography Signals using Improved Convergent Cross Mapping with Gaussian Processes
  560. Discrete Wasserstein Autoencoders for Document Retrieval
  561. Discriminant Generative Adversarial Networks with its Application to Equipment Health Classification
  562. Discriminant and Sparsity Based Least Squares Regression with l1 Regularization for Feature Representation
  563. Disentangled Multidimensional Metric Learning for Music Similarity
  564. Disentangled Speech Embeddings Using Cross-Modal Self-Supervision
  565. Disentangling Controllable Object Through Video Prediction Improves Visual Reinforcement Learning
  566. Disentangling Timbre and Singing Style with Multi-Singer Singing Synthesis System
  567. Dispersive Grid-free Orthogonal Matching Pursuit for Modal Estimation in Ocean Acoustics
  568. Distilling Attention Weights for CTC-Based ASR Systems
  569. Distributed Detection of Sparse Signals with 1-Bit Data in Two-Level Two-Degree Tree-Structured Sensor Networks
  570. Distributed Equalization and Power Allocation For Multi-Carrier Bidirectional Filter-and-Forward Relay Networks
  571. Distributed Non-Orthogonal Pilot Design for Multi-Cell Massive Mimo Systems
  572. Distributed Quantization for Sparse Time Sequences
  573. Distributed Tensor Completion Over Networks
  574. Distributed Tracking and Circumnavigation Using Bearing Measurements
  575. Distributed Verification of Belief Precisions Convergence in Gaussian Belief Propagation
  576. Distributed Wave-Domain Active Noise Control Based on the Diffusion Strategy
  577. Distribution of the Product of a Complex Gaussian Matrix and Vector and Its Sum with a Complex Gaussian Vector
  578. Divergence-Based Adaptive Extreme Video Completion
  579. Diversity and Sparsity: A New Perspective on Index Tracking
  580. Domain Adaptation for Generalization of Face Presentation Attack Detection in Mobile Settengs with Minimal Information
  581. Domain Robust, Fast, and Compact Neural Language Models
  582. Drift Detection and Correction Post-Tracking
  583. Drss-Based Localisation Using Weighted Instrumental Variables and Selective Power Measurement
  584. Dual-Path RNN: Efficient Long Sequence Modeling for Time-Domain Single-Channel Speech Separation
  585. Duration Robust Weakly Supervised Sound Event Detection
  586. Dyna-Bolt: Domain Adaptive Binary Factorization Of Current Waveforms For Energy Disaggregation
  587. Dynamic Attack Scoring Using Distributed Local Detectors
  588. Dynamic Channel Pruning For Correlation Filter Based Object Tracking
  589. Dynamic Metasurface Antennas for Bit-Constrained MIMO-OFDM Receivers
  590. Dynamic Oversampling in 1-Bit Quantized Asynchronous Large-Scale Multiple-Antenna Systems for Sustainable Iot Networks
  591. Dynamic Resource Allocation for Wireless Edge Machine Learning with Latency And Accuracy Guarantees
  592. Dynamic Resource Optimization and Altitude Selection in Uav-Based Multi-Access Edge Computing
  593. Dynamic Temporal Residual Learning for Speech Recognition
  594. Dynamic Variational Autoencoders for Visual Process Modeling
  595. Dynamically Modulated Deep Metric Learning for Visual Search
  596. Dysarthric Speech Recognition with Lattice-Free MMI
  597. E2E-SINCNET: Toward Fully End-To-End Speech Recognition
  598. ECG Heartbeat Classification Based on Multi-Scale Wavelet Convolutional Neural Networks
  599. EDNFC-Net: Convolutional Neural Network with Nested Feature Concatenation for Nuclei-Instance Segmentation
  600. EMET: Embeddings from Multilingual-Encoder Transformer for Fake News Detection
  601. EPI-Neighborhood Distribution Based Light Field Depth Estimation
  602. ESRGAN+ : Further Improving Enhanced Super-Resolution Generative Adversarial Network
  603. Edgefool: an Adversarial Image Enhancement Filter
  604. Eeg Connectivity - Informed Cooperative Adaptive Line Enhancer for Recognition of Brain State
  605. Eeg Feature Selection Using Orthogonal Regression: Application to Emotion Recognition
  606. Effect of Choice of Probability Distribution, Randomness, and Search Methods for Alignment Modeling in Sequence-to-Sequence Text-to-Speech Synthesis Using Hard Alignment
  607. Effect of Frication Duration and Formant Transitions on the Perception of Fricatives in VCV Utterances
  608. Effect of Undersampling on Non-Negative Blind Deconvolution with Autoregressive Filters
  609. Effective Approximate Maximum Likelihood Estimation of Angles of Arrival for Non-Coherent Sub-Arrays
  610. Effective Approximation of Bandlimited Signals and Their Samples
  611. Effective Pipeline for Compressing Deep Object Detectors
  612. Effective Wavenet Adaptation for Voice Conversion with Limited Data
  613. Effectiveness of Random Deep Feature Selection for Securing Image Manipulation Detectors Against Adversarial Examples
  614. Effectiveness of Self-Supervised Pre-Training for ASR
  615. Effects of Spectral Tilt on Listeners' Preferences And Intelligibility
  616. Efficient Algorithm to Implement Sliding Singular Spectrum Analysis with Application to Biomedical Signal Denoising
  617. Efficient Belief Propagation for Graph Matching
  618. Efficient Bird Sound Detection on the Bela Embedded System
  619. Efficient Constrained Encoders Correcting a Single Nucleotide Edit in DNA Storage
  620. Efficient Decoupled Neural Architecture Search by Structure And Operation Sampling
  621. Efficient Deep Learning-Based Lossy Image Compression Via Asymmetric Autoencoder and Pruning
  622. Efficient Estimation of Mixing Matrix Using a Two-sensor Array
  623. Efficient Image Super Resolution Via Channel Discriminative Deep Neural Network Pruning
  624. Efficient Multichannel Nonlinear Acoustic Echo Cancellation Based on a Cooperative Strategy
  625. Efficient Scene Text Detection with Textual Attention Tower
  626. Efficient Shallow Wavenet Vocoder Using Multiple Samples Output Based on Laplacian Distribution and Linear Prediction
  627. Efficient Super-Resolution Two-Dimensional Harmonic Retrieval Via Enhanced Low-Rank Structured Covariance Reconstruction
  628. Efficient Techniques For in-band System Information Broadcast in Multi-Cell Massive Mimo
  629. Efficient Trainable Front-Ends for Neural Speech Enhancement
  630. Efficient and Scalable Neural Residual Waveform Coding with Collaborative Quantization
  631. Electric Analog Circuit Design with Hypernetworks And A Differential Simulator
  632. Electro-Magnetic Side-Channel Attack Through Learned Denoising and Classification
  633. Eliminating Out-Of-Cell Interference in Cellular Massive Mimo with a Single Additional Transceiver
  634. Embedded Large-Scale Handwritten Chinese Character Recognition
  635. Emotional Speech Synthesis with Rich and Granularized Control
  636. Emotional Voice Conversion Using Multitask Learning with Text-To-Speech
  637. Empirical Sure-Guided Microscopy Super-Resolution Image Reconstruction from Confocal Multi-Array Detectors
  638. Encoder-Recurrent Decoder Network for Single Image Dehazing
  639. Encoding Temporal Information For Automatic Depression Recognition From Facial Analysis
  640. Encoding and Decoding Mixed Bandlimited Signals Using Spiking Integrate-and-Fire Neurons
  641. End to End Speech Recognition Error Prediction with Sequence to Sequence Learning
  642. End-To-End Accent Conversion Without Using Native Utterances
  643. End-To-End Auditory Object Recognition Via Inception Nucleus
  644. End-To-End Generation of Talking Faces from Noisy Speech
  645. End-To-End Multi-Speaker Speech Recognition With Transformer
  646. End-To-End Multi-Talker Overlapping Speech Recognition
  647. End-To-End Non-Negative Autoencoders for Sound Source Separation
  648. End-To-End Spoken Language Understanding Without Matched Language Speech Model Pretraining Data
  649. End-To-End Voice Conversion Via Cross-Modal Knowledge Distillation for Dysarthric Speech Reconstruction
  650. End-end Speech-to-Text Translation with Modality Agnostic Meta-Learning
  651. End-to-End Architectures for ASR-Free Spoken Language Understanding
  652. End-to-End Articulatory Modeling for Dysarthric Articulatory Attribute Detection
  653. End-to-End Automatic Speech Recognition Integrated with CTC-Based Voice Activity Detection
  654. End-to-End Code-Switching TTS with Cross-Lingual Language Model
  655. End-to-End Multi-Person Audio/Visual Automatic Speech Recognition
  656. End-to-End Speech Translation with Self-Contained Vocabulary Manipulation
  657. End-to-End Training of Time Domain Audio Separation and Recognition
  658. End-to-end Microphone Permutation and Number Invariant Multi-channel Speech Separation
  659. EnerGAN: A GENERATIVE ADVERSARIAL NETWORK FOR ENERGY DISAGGREGATION
  660. Energy Disaggregation Using Fractional Calculus
  661. Energy Disaggregation from Low Sampling Frequency Measurements Using Multi-Layer Zero Crossing Rate
  662. Energy Efficient Acceleration Of Floating Point Applications Onto CGRA
  663. Energy-Efficient 3D UAV Trajectory Design for Data Collection in Wireless Sensor Networks
  664. Energy-Efficient Bit Allocation for Resolution-Adaptive ADC in Multiuser Large-Scale MIMO Systems: Global Optimality
  665. Enhance Part-Based Model for Person Re-Identification with Fused Multi-Scale Features
  666. Enhance feature representation of electroencephalogram for Seizure detection
  667. Enhanced Action Tubelet Detector for Spatio-Temporal Video Action Detection
  668. Enhanced Adversarial Strategically-Timed Attacks Against Deep Reinforcement Learning
  669. Enhanced Method of Audio Coding Using CNN-Based Spectral Recovery with Adaptive Structure
  670. Enhanced Mixture Population Monte Carlo Via Stochastic Optimization and Markov Chain Monte Carlo Sampling
  671. Enhanced Non-Local Cascading Network with Attention Mechanism for Hyperspectral Image Denoising
  672. Enhanced Safety of Autonomous Driving by Incorporating Terrestrial Signals of Opportunity
  673. Enhancement of Coded Speech Using a Mask-Based Post-Filter
  674. Enhancing End-to-End Multi-Channel Speech Separation Via Spatial Feature Learning
  675. Enhancing the Labelling of Audio Samples for Automatic Instrument Classification Based on Neural Networks
  676. Ensemble Network For Ranking Images Based On Visual Appeal
  677. Environment-Aware Reconfigurable Noise Suppression
  678. Epigraphical Reformulation for Non-Proximable Mixed Norms
  679. Epoch Estimation from a Speech Signal Using Gammatone Wavelets in a Scattering Network
  680. Equalization of OFDM Waveforms with Insufficient Cyclic Prefix
  681. Ernet Family: Hardware-Oriented Cnn Models For Computational Imaging Using Block-Based Inference
  682. Error Analysis Applied to End-to-End Spoken Language Understanding
  683. Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to-Speech Toolkit
  684. Estimating Centrality Blindly From Low-Pass Filtered Graph Signals
  685. Estimating Structural Missing Values Via Low-Tubal-Rank Tensor Completion
  686. Estimating the Degree of Sleepiness by Integrating Articulatory Feature Knowledge in Raw Waveform Based CNNS
  687. Estimation of Information in Parallel Gaussian Channels via Model Order Selection
  688. Estimation of Post-Nonlinear Causal Models Using Autoencoding Structure
  689. Europarl-ST: A Multilingual Corpus for Speech Translation of Parliamentary Debates
  690. Evaluating Voice Conversion-Based Privacy Protection against Informed Attackers
  691. Evaluation of Deep-Learning-Based Voice Activity Detectors and Room Impulse Response Models in Reverberant Environments
  692. Evaluation of Joint Auditory Attention Decoding and Adaptive Binaural Beamforming Approach for Hearing Devices with Attention Switching
  693. Evaluation of Sensor Self-Noise In Binaural Rendering of Spherical Microphone Array Signals
  694. Event-Driven Signal Processing with Neuromorphic Computing Systems
  695. Exact Sparse Nonnegative Least Squares
  696. Exocentric to Egocentric Image Generation Via Parallel Generative Adversarial Network
  697. Experiments in Creating Online Course Content for Signal Processing Education
  698. Exploitation of 3D City Maps for Hybrid 5G RTT and GNSS Positioning Simulations
  699. Exploiting Channel Locality for Adaptive Massive MIMO Signal Detection
  700. Exploiting Commutativity Condition for CP Decomposition Via Approximate Simultaneous Diagonalization
  701. Exploiting Periodicity Features for Joint Detection and DOA Estimation of Speech Sources Using Convolutional Neural Networks
  702. Exploiting Rays in Blind Localization of Distributed Sensor Arrays
  703. Exploiting Sparsity for Robust Sensor Network Localization in Mixed LOS/NLOS Environments
  704. Exploiting Two-Dimensional Symmetry and Unimodality for Model-Free Source Localization in Harsh Environment
  705. Exploiting Vocal Tract Coordination Using Dilated CNNS For Depression Detection In Naturalistic Environments
  706. Exploration Methodology for BTI-Induced Failures on RRAM-Based Edge AI Systems
  707. Exploring A Zero-Order Direct Hmm Based on Latent Attention for Automatic Speech Recognition
  708. Exploring Appropriate Acoustic and Language Modelling Choices for Continuous Dysarthric Speech Recognition
  709. Exploring Bio-Behavioral Signal Trajectories of State Anxiety During Public Speaking
  710. Exploring Energy Efficient Quantum-resistant Signal Processing Using Array Processors
  711. Exploring Entity-Level Spatial Relationships for Image-Text Matching
  712. Exploring Pre-Training with Alignments for RNN Transducer Based End-to-End Speech Recognition
  713. Exposure Interpolation Via Hybrid Learning
  714. Expression-Guided EEG Representation Learning for Emotion Recognition
  715. Extended Cyclic Coordinate Descent for Robust Row-Sparse Signal Reconstruction in the Presence of Outliers
  716. Extended Object Tracking Using Hierarchical Truncation Measurement Model with Automotive Radar
  717. Extracting Unit Embeddings Using Sequence-To-Sequence Acoustic Models for Unit Selection Speech Synthesis
  718. Extrapolated Alternating Algorithms for Approximate Canonical Polyadic Decomposition
  719. F0-Consistent Many-To-Many Non-Parallel Voice Conversion Via Conditional Autoencoder
  720. FCEM: A Novel Fast Correlation Extract Model For Real Time Steganalysis Of VoIP Stream Via Multi-Head Attention
  721. FDDWNet: A Lightweight Convolutional Neural Network for Real-Time Semantic Segmentation
  722. FIR Filter Design and Implementation for Phase-Based Processing
  723. FIR Filtering of Discontinuous Signals: A Random-Stratified Sampling Approach
  724. Face Feature Recovery via Temporal Fusion for Person Search
  725. Facial Emotion Recognition Using Light Field Images with Deep Attention-Based Bidirectional LSTM
  726. Facial Feature Embedded Cyclegan For Vis-Nir Translation
  727. Far-Field Location Guided Target Speech Extraction Using End-to-End Speech Recognition Objectives
  728. Fast Acoustic Scattering Using Convolutional Neural Networks
  729. Fast Block-Sparse Estimation for Vector Networks
  730. Fast Clustering With Co-Clustering Via Discrete Non-Negative Matrix Factorization for Image Identification
  731. Fast Direction-of-arrival Estimation of Multiple Targets Using Deep Learning and Sparse Arrays
  732. Fast Domain Adaptation for Goal-Oriented Dialogue Using a Hybrid Generative-Retrieval Transformer
  733. Fast Independent Vector Extraction by Iterative SINR Maximization
  734. Fast Intent Classification for Spoken Language Understanding Systems
  735. Fast Lattice-Free Keyword Filtering for Accelerated Spoken Term Detection
  736. Fast Optical System Identification by Numerical Interferometry
  737. Fast Single-View 3D Object Reconstruction with Fine Details Through Dilated Downsample and Multi-Path Upsample Deep Neural Network
  738. Fast Start-Up Algorithm for Adaptive Noise Cancellers with Novel SNR Estimation and Stepsize Control
  739. Fast Training of Deep Neural Networks for Speech Recognition
  740. Fast and Accurate Embedded DCNN for Rgb-D Based Sign Language Recognition
  741. Fast and High-Quality Singing Voice Synthesis System Based on Convolutional Neural Networks
  742. Fast and Stable Blind Source Separation with Rank-1 Updates
  743. Faster-Than-Nyquist Signaling Via Spatiotemporal Symbol-Level Precoding for Multi-User MISO Redundant Transmissions
  744. Favorable Propagation and Linear Multiuser Detection for Distributed Antenna Systems
  745. Feature Affine Projection Algorithms
  746. Feature Drift Resilient Tracking of The Carotid Artery Wall Using Unscented Kalman Filtering With Data Fusion
  747. Feature Enhancement with Deep Feature Losses for Speaker Verification
  748. Feature Selection Under Orthogonal Regression with Redundancy Minimizing
  749. Federated Classification with Low Complexity Reproducing Kernel Hilbert Space Representations
  750. Federated Learning with Mutually Cooperating Devices: A Consensus Approach Towards Server-Less Model Optimization
  751. Federated Learning with Quantization Constraints
  752. Federated Neuromorphic Learning of Spiking Neural Networks for Low-Power Edge Intelligence
  753. Federated Truth Inference over Distributed Crowdsourcing Platforms
  754. Federating Solar, Storage and Communications in the Electric Grid and Internet of things
  755. Feedback Recurrent Autoencoder
  756. Feedback Turbo Autoencoder
  757. Few-Shot Acoustic Event Detection Via Meta Learning
  758. Few-Shot Sound Event Detection
  759. Fg2seq: Effectively Encoding Knowledge for End-To-End Task-Oriented Dialog
  760. Filterbank Design for End-to-end Speech Separation
  761. Filtering Out Time-Frequency Areas Using Gabor Multipliers
  762. Fine-Grained Action Recognition on a Novel Basketball Dataset
  763. Fine-Grained Giant Panda Identification
  764. Finite Sample Deviation and Variance Bounds for First Order Autoregressive Processes
  765. Fixed Smooth Convolutional Layer for Avoiding Checkerboard Artifacts in CNNS
  766. Fixed-Point Optimization of Transformer Neural Network
  767. Flexibly-tunable bitcube-based perceptual encryption within jpeg compression
  768. Flow-TTS: A Non-Autoregressive Network for Text to Speech Based on Flow
  769. Focus on Semantic Consistency for Cross-Domain Crowd Understanding
  770. Focusing on Attention: Prosody Transfer and Adaptative Optimization Strategy for Multi-Speaker End-to-End Speech Synthesis
  771. Forecasting Multi-Dimensional Processes Over Graphs
  772. Forecasting Sparse Traffic Congestion Patterns Using Message-Passing RNNS
  773. Foreground Signature Extraction for an Intimate Mixing Model in Hyperspectral Image Classification
  774. Formulating Divergence Framework for Multiclass Motor Imagery EEG Brain Computer Interface
  775. Forward-Backward Splitting for Optimal Transport Based Problems
  776. Fourier Phase Retrieval with Arbitrary Reference Signal
  777. Fourth Order Cumulant Based Active Direction of Arrival Estimation Using Coprime Arrays
  778. Fractional Fourier Transform Based QRS Complex Detection in ECG Signal
  779. Frame-Based Overlapping Speech Detection Using Convolutional Neural Networks
  780. Frame-Level MMI as A Sequence Discriminative Training Criterion for LVCSR
  781. Frame-Level Phoneme-Invariant Speaker Embedding for Text-Independent Speaker Recognition on Extremely Short Utterances
  782. Frequency Diverse Array Radar: A Closed-Form Solution to Design Weights for Desired Beampattern
  783. Frequency and Temporal Convolutional Attention for Text-Independent Speaker Recognition
  784. Frequency-Dependent Directional Feedback Delay Network
  785. From Symbols to Signals: Symbolic Variational Autoencoders
  786. From Unsupervised Machine Translation to Adversarial Text Generation
  787. From Video Game to Real Robot: The Transfer Between Action Spaces
  788. Full Reference Video Quality Measures Improvement Using Neural Networks
  789. Full-Reference Speech Quality Estimation with Attentional Siamese Neural Networks
  790. Full-Sum Decoding for Hybrid Hmm Based Speech Recognition Using LSTM Language Model
  791. Fully Convolutional Recurrent Networks for Speech Enhancement
  792. Fully Learnable Front-End for Multi-Channel Acoustic Modeling Using Semi-Supervised Learning
  793. Fully Pipelined Iteration Unrolled Decoders the Road to TB/S Turbo Decoding
  794. Fully-Hierarchical Fine-Grained Prosody Modeling For Interpretable Speech Synthesis
  795. Fully-Neural Approach to Heavy Vehicle Detection on Bridges Using a Single Strain Sensor
  796. Fusion Approaches for Emotion Recognition from Speech Using Acoustic and Text-Based Features
  797. Fusionndvi: A Novel Fusion Method for NDVI in Remote Sensing
  798. G2G: TTS-Driven Pronunciation Learning for Graphemic Hybrid ASR
  799. GCI Detection from Raw Speech Using a Fully-Convolutional Network
  800. GFCN: A New Graph Convolutional Network Based on Parallel Flows
  801. GFNet: A Lightweight Group Frame Network for Efficient Human Action Recognition
  802. Gait Phase Segmentation Using Weighted Dynamic Time Warping and K-Nearest Neighbors Graph Embedding
  803. Gated Attentive Convolutional Network Dialogue State Tracker
  804. Gated Mechanism for Attention Based Multi Modal Sentiment Analysis
  805. Gated Multi-Layer Convolutional Feature Extraction Network for Robust Pedestrian Detection
  806. Gaussian Lpcnet for Multisample Speech Synthesis
  807. Gaussian Process Imputation of Multiple Financial Series
  808. Gaussian Processes Over Graphs
  809. Gender Differences on the Perception and Production of Utterances with Willingness and Reluctance in Chinese
  810. Generalized Coherence-Based Signal Enhancement
  811. Generalized Graph Spectral Sampling with Stochastic Priors
  812. Generalized Kernel-Based Dynamic Mode Decomposition
  813. Generalized Linear Bandits with Safety Constraints
  814. Generalized Spatial Modulation for Wireless Terabits Systems Under Sub-THZ Channel With RF Impairments
  815. Generating Diverse and Natural Text-to-Speech Samples Using a Quantized Fine-Grained VAE and Autoregressive Prosody Prior
  816. Generating Empathetic Responses by Looking Ahead the User's Sentiment
  817. Generating Multilingual Voices Using Speaker Space Translation Based on Bilingual Speaker Data
  818. Generating Synthetic Audio Data for Attention-Based Speech Recognition Systems
  819. Generating and Protecting Against Adversarial Attacks for Deep Speech-Based Emotion Recognition Models
  820. Generative Adversarial Networks for Graph Data Imputation from Signed Observations
  821. Generative Pre-Training for Speech with Autoregressive Predictive Coding
  822. Genetic Algorithm Optimized Support Vector Machine in NOMA-based Satellite Networks with Imperfect CSI
  823. Geometrically Constrained Independent Vector Analysis for Directional Speech Enhancement
  824. Geometry Constrained Progressive Learning for Lstm-Based Speech Enhancement
  825. Global Structure Graph Guided Fine-Grained Vehicle Recognition
  826. Global Traffic State Recovery VIA Local Observations with Generative Adversarial Networks
  827. Global and Local Discriminative Patches Exploiting for Action Recognition
  828. Gpu-Accelerated Viterbi Exact Lattice Decoder for Batched Online and Offline Speech Recognition
  829. Gradient Delay Analysis in Asynchronous Distributed Optimization
  830. Gradient-Based Algorithm with Spatial Regularization for Optimal Sensor Placement
  831. Graph Auto-Encoder for Graph Signal Denoising
  832. Graph Construction from Data by Non-Negative Kernel Regression
  833. Graph Convolutional Neural Networks to Classify Whole Slide Images
  834. Graph Metric Learning via Gershgorin Disc Alignment
  835. Graph Neural Net Using Analytical Graph Filters and Topology Optimization for Image Denoising
  836. Graph Vertex Sampling with Arbitrary Graph Signal Hilbert Spaces
  837. GraphTTS: Graph-to-Sequence Modelling in Neural Text-to-Speech
  838. Graphem: EM Algorithm for Blind Kalman Filtering Under Graphical Sparsity Constraints
  839. Graphical Evolutionary Game Theoretic Analysis of Super Users in Information Diffusion
  840. Gray-Scale Image Colorization Using Cycle-Consistent Generative Adversarial Networks with Residual Structure Enhancer
  841. Greedy Hybrid Rate Adaptation in Dynamic Wireless Communication Environment
  842. Greedy Sparse Array Design for Optimal Localization under Spatially Prioritized Source Distribution
  843. Group-Utility Metric for Efficient Sensor Selection and Removal in LCMV Beamformers
  844. Guided Learning for Weakly-Labeled Semi-Supervised Sound Event Detection
  845. Gyroscope Aided Video Stabilization Using Nonlinear Regression on Special Orthogonal Group
  846. H-Vectors: Utterance-Level Speaker Embedding Using a Hierarchical Attention Model
  847. HDMFH: Hypergraph Based Discrete Matrix Factorization Hashing for Multimodal Retrieval
  848. HGFM : A Hierarchical Grained and Feature Model for Acoustic Emotion Recognition
  849. HI-MIA: A Far-Field Text-Dependent Speaker Verification Database and the Baselines
  850. HKA: A Hierarchical Knowledge Attention Mechanism for Multi-Turn Dialogue System
  851. HPRNN: A Hierarchical Sequence Prediction Model for Long-Term Weather Radar Echo Extrapolation
  852. Hand-3d-Studio: A New Multi-View System for 3d Hand Reconstruction
  853. Harmonic/Percussive Sound Separation and Spectral Complexity Reduction of Music Signals for Cochlear Implant Listeners
  854. Harmonics Based Representation in Clarinet Tone Quality Evaluation
  855. Headless Horseman: Adversarial Attacks on Transfer Learning Models
  856. Hearing aid Research Data Set for Acoustic Environment Recognition
  857. Height and Weight Estimation from Unconstrained Images
  858. Heterogeneous Domain Generalization Via Domain Mixup
  859. Hidden Markov Models for Sepsis Detection in Preterm Infants
  860. Hierarchical Attention Transfer Networks for Depression Assessment from Speech
  861. Hierarchical Caching via Deep Reinforcement Learning
  862. Hierarchical Federated Learning ACROSS Heterogeneous Cellular Networks
  863. Hierarchical Sequence Representation with Graph Network
  864. High Dynamic Range Imaging Using Deep Image Priors
  865. High-Accuracy Classification of Attention Deficit Hyperactivity Disorder with L2, 1-Norm Linear Discriminant Analysis
  866. High-Accuracy and Low-Latency Speech Recognition with Two-Head Contextual Layer Trajectory LSTM Model
  867. High-Dimensional Neural Feature Using Rectified Linear Unit And Random Matrix Instance
  868. High-Resolution Attention Network with Acoustic Segment Model for Acoustic Scene Classification
  869. Hijacking Tracker: A Powerful Adversarial Attack on Visual Tracking
  870. How Much Self-Attention Do We Need? Trading Attention for Feed-Forward Layers
  871. How confident are you? Exploring the role of fillers in the automatic prediction of a speaker's confidence
  872. Human-Machine Collaboration for Medical Image Segmentation
  873. Humangan: Generative Adversarial Network With Human-Based Discriminator And Its Evaluation In Speech Perception Modeling
  874. Humbug Zooniverse: A Crowd-Sourced Acoustic Mosquito Dataset
  875. Hybrid Active Contour Driven by Double-Weighted Signed Pressure Force for Image Segmentation
  876. Hybrid Autoregressive Transducer (HAT)
  877. Hybrid Deep-Semantic Matrix Factorization for Tag-Aware Personalized Recommendation
  878. Hybrid Neural-Parametric F0 Model for Singing Synthesis
  879. Hybrid Precoding for Secure Transmission in Reflect-Array-Assisted Massive MIMO Systems
  880. Hydranet: A Real-Time Waveform Separation Network
  881. I-Vector Transformation Using K-Nearest Neighbors for Speaker Verification
  882. IQ-STAN: Image Quality Guided Spatio-Temporal Attention Network for License Plate Recognition
  883. Identification of Essential Proteins Using A Novel Multi-Objective Optimization Method
  884. Identifying Truthful Language in Child Interviews
  885. Image Fusion using Joint Sparse Representations and Coupled Dictionary Learning
  886. Image Processing in DNA
  887. Image Recovery from Rotational And Translational Invariants
  888. Image Restoration Via Data-Dependent Proximal Averaged Optimization
  889. Image Segmentation Based Privacy-Preserving Human Action Recognition for Anomaly Detection
  890. Image Super-Resolution Using Residual Global Context Network
  891. Impact of a Shift-Invariant Harmonic Phase Model in Fully Parametric Harmonic Voice Representation and Time/Frequency Synthesis
  892. Improved End-To-End Spoken Utterance Classification with a Self-Attention Acoustic Classifier
  893. Improved Large-Margin Softmax Loss for Speaker Diarisation
  894. Improved Nearest Neighbor Density-Based Clustering Techniques with Application to Hyperspectral Images
  895. Improved Probability Modelling for Exception Handling in Lossless Screen Content Coding
  896. Improved Real-Time Visual Tracking via Adversarial Learning
  897. Improved Speaker Independent Dysarthria Intelligibility Classification Using Deepspeech Posteriors
  898. Improving Auditory Attention Decoding Performance of Linear and Non-Linear Methods using State-Space Model
  899. Improving Automated Segmentation of Radio Shows with Audio Embeddings
  900. Improving Convergent Cross Mapping for Causal Discovery with Gaussian Processes
  901. Improving Cross-Dataset Performance of Face Presentation Attack Detection Systems Using Face Recognition Datasets
  902. Improving Deep CNN Networks with Long Temporal Context for Text-Independent Speaker Verification
  903. Improving Deep Learning Classification of JPEG2000 Images Over Bandlimited Networks
  904. Improving Device Directedness Classification of Utterances With Semantic Lexical Features
  905. Improving Efficiency in Large-Scale Decentralized Distributed Training
  906. Improving End-to-End Speech Synthesis with Local Recurrent Neural Network Enhanced Transformer
  907. Improving LPCNET-Based Text-to-Speech with Linear Prediction-Structured Mixture Density Network
  908. Improving Language Identification for Multilingual Speakers
  909. Improving Music Transcription by Pre-Stacking A U-Net
  910. Improving Noise Robust Automatic Speech Recognition with Single-Channel Time-Domain Enhancement Network
  911. Improving Proper Noun Recognition in End-To-End Asr by Customization of the Mwer Loss Criterion
  912. Improving Prosody with Linguistic and Bert Derived Features in Multi-Speaker Based Mandarin Chinese Neural TTS
  913. Improving Reverberant Speech Training Using Diffuse Acoustic Simulation
  914. Improving Robustness of Deep Learning Based Monaural Speech Enhancement Against Processing Artifacts
  915. Improving Sample-Efficiency in Reinforcement Learning for Dialogue Systems by Using Trainable-Action-Mask
  916. Improving Sequence-To-Sequence Speech Recognition Training with On-The-Fly Data Augmentation
  917. Improving Singing Voice Separation with the Wave-U-Net Using Minimum Hyperspherical Energy
  918. Improving Speaker Discrimination of Target Speech Extraction With Time-Domain Speakerbeam
  919. Improving Speaker-Attribute Estimation by Voting Based on Speaker Cluster Information
  920. Improving Speech Recognition Using Consistent Predictions on Synthesized Speech
  921. Improving Spoken Question Answering Using Contextualized Word Representation
  922. Improving Universal Sound Separation Using Sound Classification
  923. Improving Voice Separation by Incorporating End-To-End Speech Recognition
  924. Improving the Chronological Sorting of Images through Occlusion: A Study on the Notre-Dame Cathedral Fire
  925. Improving the Performance of Transformer Based Low Resource Speech Recognition for Indian Languages
  926. Improving the Scalability of Deep Reinforcement Learning-Based Routing with Control on Partial Nodes
  927. Impulse Response Data Augmentation and Deep Neural Networks for Blind Room Acoustic Parameter Estimation
  928. In-Domain and Out-of-Domain Data Augmentation to Improve Children's Speaker Verification System in Limited Data Scenario
  929. In-Network Caching for Hybrid Satellite-Terrestrial Networks Using Deep Reinforcement Learning
  930. Incorporating Written Domain Numeric Grammars into End-To-End Contextual Speech Recognition Systems for Improved Recognition of Numeric Sequences
  931. Incremental Semi-Supervised Learning for Multi-Genre Speech Recognition
  932. Independent Language Modeling Architecture for End-To-End ASR
  933. Individual Distance-Dependent HRTFS Modeling Through A Few Anthropometric Measurements
  934. Indoor Altitude Estimation of Unmanned Aerial Vehicles Using a Bank of Kalman Filters
  935. Indoor Heading Direction Estimation Using Rf Signals
  936. Indylstms: Independently Recurrent LSTMS
  937. Inferring Dynamic Group Leadership Using Sequential Bayesian Methods
  938. Information Flow Optimization in Inference Networks
  939. Information Maximized Variational Domain Adversarial Learning for Speaker Verification
  940. Information Theoretic Approach for Waveform Design in Coexisting MIMO Radar and MIMO Communications
  941. Instance-based Model Adaptation for Direct Speech Translation
  942. Instant Adaptive Learning: An Adaptive Filter Based Fast Learning Model Construction for Sensor Signal Time Series Classification on Edge Devices
  943. Integrating Discrete and Neural Features Via Mixed-Feature Trans-Dimensional Random Field Language Models
  944. Integration of Multi-Look Beamformers for Multi-Channel Keyword Spotting
  945. Intelligent Reflecting Surface for Massive Device Connectivity: Joint Activity Detection and Channel Estimation
  946. Intelligent Student Behavior Analysis System for Real Classrooms
  947. Intensity-Image Reconstruction for Event Cameras Using Convolutional Neural Network
  948. Interpolation and Range Extrapolation of Sound Source Directivity Based on a Spherical Wave Propagation Model
  949. Interpretability-Guided Convolutional Neural Networks for Seismic Fault Segmentation
  950. Interpretable Machine Learning In Sustainable Edge Computing: A Case Study of Short-Term Photovoltaic Power Output Prediction
  951. Interpretable Self-Attention Temporal Reasoning for Driving Behavior Understanding
  952. Interrupted and Cascaded Permutation Invariant Training for Speech Separation
  953. Intra Frame Rate Control for Versatile Video Coding with Quadratic Rate-Distortion Modelling
  954. Inverse Multiple Scattering with Phaseless Measurements
  955. Invertible DNN-Based Nonlinear Time-Frequency Transform for Speech Enhancement
  956. Investigating Generalization in Neural Networks Under Optimally Evolved Training Perturbations
  957. Investigation of Methods to Improve the Recognition Performance of Tamil-English Code-Switched Data in Transformer Framework
  958. Investigation of Specaugment for Deep Speaker Embedding Learning
  959. JPEG Steganography with Side Information from the Processing Pipeline
  960. Jhu-HLTCOE System for the Voxsrc Speaker Recognition Challenge
  961. Joint Beamforming and Reverberation Cancellation Using a Constrained Kalman Filter With Multichannel Linear Prediction
  962. Joint Blind Calibration and Time-Delay Estimation for Multiband Ranging
  963. Joint Coding and Modulation in the Ultra-Short Blocklength Regime for Bernoulli-Gaussian Impulsive Noise Channels Using Autoencoders
  964. Joint Contextual Modeling for ASR Correction and Language Understanding
  965. Joint Enhancement And Denoising of Low Light Images Via JND Transform
  966. Joint Estimation Of Acoustic Parameters From Single-Microphone Speech Observations
  967. Joint Frequency Domain Channel Estimation and Equalization Based on Expectation Propagation for Single Carrier Transmissions
  968. Joint Learning of Assignment and Representation for Biometric Group Membership
  969. Joint Learning of Cartesian under Sampling Andre Construction for Accelerated MRI
  970. Joint Multitarget Tracking and Dynamic Network Localization in the Underwater Domain
  971. Joint Optimization of Sampling Patterns and Deep Priors for Improved Parallel MRI
  972. Joint Phoneme Alignment and Text-Informed Speech Separation on Highly Corrupted Speech
  973. Joint Phoneme-Grapheme Model for End-To-End Speech Recognition
  974. Joint Resource Allocation and Routing for Service Function Chaining with In-Subnetwork Processing
  975. Joint Scheduling and Beamforming for Delay Sensitive Traffic with Priorities and Deadlines
  976. Joint Semi-Supervised Feature Auto-Weighting and Classification Model for EEG-Based Cross-Subject Sleep Quality Evaluation
  977. Joint Source-Channel Coding and Bayesian Message Passing Detection for Grant-Free Radio Access in IoT
  978. Joint Sparse Recovery Using Deep Unfolding With Application to Massive Random Access
  979. Joint Training of Deep Neural Networks for Multi-Channel Dereverberation and Speech Source Separation
  980. Jointly Optimal Dereverberation and Beamforming
  981. Just Noticeable Distortion Based Perceptually Lossless Intra Coding
  982. K-Autoencoders Deep Clustering
  983. K-Space Trajectory Design for Reduced MRI Scan Time
  984. KALM: Key Area Localization Mechanism for Abnormality Detection in Musculoskeletal Radiographs
  985. Kernel Computations from Large-Scale Random Features Obtained by Optical Processing Units
  986. Kernel Ridge Regression with Autocorrelation Prior: Optimal Model and Cross-Validation
  987. Key Action and Joint CTC-Attention based Sign Language Recognition
  988. Keyword Search for Sign Language
  989. Knowledge Distillation and Random Erasing Data Augmentation for Text-Dependent Speaker Verification
  990. Knowledge Enhanced Latent Relevance Mining for Question Answering
  991. Korean Singing Voice Synthesis Based on Auto-Regressive Boundary Equilibrium Gan
  992. L-Vector: Neural Label Embedding for Domain Adaptation
  993. L1-Norm Higher-Order Orthogonal Iterations for Robust Tensor Analysis
  994. LEt-SNE: A Hybrid Approach to Data Embedding and Visualization Of Hyperspectral Imagery
  995. LSTM-Based One-Pass Decoder for Low-Latency Streaming
  996. Label Propagation Adaptive Resonance Theory for Semi-Supervised Continuous Learning
  997. Label Reuse for Efficient Semi-Supervised Learning
  998. Lai-Net: Local-Ancestry Inference with Neural Networks
  999. Lance: efficient low-precision quantized winograd convolution for neural networks based on graphics processing units
  1000. Language Independent Gender Identification from Raw Waveform Using Multi-Scale Convolutional Neural Networks

Looking for submission deadlines instead? See the conference deadline calendar.