Other

Quantum Federated Learning with Quantum Networks PPT

Read more about Quantum Federated Learning with Quantum Networks PPT
Log in to post comments

A major concern of deep learning models is the large amount of data that is required to build and train them, much of which is reliant on sensitive and personally identifiable information that is vulnerable to access by third parties. Ideas of using the quantum internet to address this issue have been previously proposed, which would enable fast and completely secure online communications. Previous work has yielded a hybrid quantum-classical transfer learning scheme for classical data and communication with a hub-spoke topology.

Quantum Federated Learning with Quantum Networks ICASSP 2024.pptx

Quantum Federated Learning with Quantum Networks ICASSP 2024.pptx (19)

Categories:: Other

10 Views

Object Trajectory Estimation with Multi-Band Wi-Fi Neural Dynamic Fusion

Read more about Object Trajectory Estimation with Multi-Band Wi-Fi Neural Dynamic Fusion
Log in to post comments

In contrast to existing multi-band Wi-Fi fusion in a frame-to-frame basis for simple classification, this paper considers asynchronous sequence-to-sequence fusion between sub-7GHz channel state information (CSI) and 60GHz beam SNR for more challenging downstream tasks such as continuous regression.

icassp_skato_release.pptx

icassp_skato_release.pptx (8)

Categories:: Other

21 Views

Poster for ICASSP 2024 paper "Turn-taking and Backchannel Prediction with Acoustic and Large Language Model Fusion"

We propose an approach for continuous prediction of turn-taking and backchanneling locations in spoken dialogue by fusing a neural acoustic model with a large language model (LLM). Experiments on the Switchboard human-human conversation dataset demonstrate that our approach consistently outperforms the baseline models with single modality. We also develop a novel multi-task instruction fine-tuning strategy to further benefit from LLM-encoded knowledge for understanding the tasks and conversational contexts, leading to additional improvements.

Turn-taking LLM ICASSP2024 Poster_v2 (1).pdf

Turn-taking LLM ICASSP2024 Poster_v2 (1).pdf (11)

Categories:: Other

7 Views

Poster for ICASSP 2024 paper "Hot-Fixing Wake Work Recognition for End-to-End ASR via Neural Model Reprogramming"

This paper proposes two novel variants of neural reprogramming to enhance wake word recognition in streaming end-to-end ASR models without updating model weights. The first, "trigger-frame reprogramming", prepends the input speech feature sequence with the learned trigger-frames of the target wake word to adjust ASR model’s hidden states for improved wake word recognition. The second, "predictor-state initialization", trains only the initial state vectors (cell and hidden states) of the LSTMs in the prediction network.

WW_HF_w_NP_ICASSP2024 Poster.pdf

WW_HF_w_NP_ICASSP2024 Poster.pdf (13)

Categories:: Other

7 Views

Improving Medical Dialogue Generation with Abstract Meaning Representations

Read more about Improving Medical Dialogue Generation with Abstract Meaning Representations
Log in to post comments

Medical Dialogue Generation plays a critical role in telemedicine by facilitating the dissemination of medical expertise to patients. Existing studies focus on incorporating textual representations, which have limited their ability to represent text semantics, such as ignoring important medical entities.

Improving__Medical_Dialogue_Generation_with_Abstract_Meaning_Representations__ICASSP_2024_.pdf

paper (8)

oral_icassp.pptx

slides (11)

Categories:: Other
Knowledge and Data Engineering
Spoken Language Processing

2 Views

Privacy Preserving Federated Learning from Multi-input Functional Proxy Re-encryption

Read more about Privacy Preserving Federated Learning from Multi-input Functional Proxy Re-encryption
Log in to post comments

Federated learning (FL) allows different participants to collaborate on model training without transmitting raw data, thereby protecting user data privacy. However, FL faces a series of security and privacy issues (e.g. the leakage of raw data from publicly shared parameters). Several privacy protection technologies, such as homomorphic encryption, differential privacy and functional encryption, are introduced for privacy enhancement in FL. Among them, the FL frameworks based on functional encryption better balance security and performance, thus receiving increasing attention.

ICASSP24_Poster___MI_FPRE (1).pdf

ICASSP__24___Poster___MI_FPRE (1).pdf (17)

Categories:: Other

7 Views

Poster for IMAGE ATTRIBUTION BY GENERATING IMAGES

Read more about Poster for IMAGE ATTRIBUTION BY GENERATING IMAGES
Log in to post comments

We introduce GPNN-CAM, a novel method for CNN explanation, that bridges two distinct areas of computer vision:
Image Attribution, which aims to explain a predictor by highlighting image regions it finds important, and Single
Image Generation (SIG), that focuses on learning how to generate variations of a single sample. GPNN-CAM leverages samples generated by Generative

ICASSP-poster.pdf

ICASSP-poster.pdf (13)

Categories:: Other

5 Views

A UNIFIED DNN-BASED SYSTEM FOR INDUSTRIAL PIPELINE SEGMENTATION

Read more about A UNIFIED DNN-BASED SYSTEM FOR INDUSTRIAL PIPELINE SEGMENTATION
Log in to post comments

This paper presents a unified system tailored for autonomous pipe segmentation within an industrial setting. To this end, it is designed to analyze RGB images captured by Unmanned Aerial Vehicle (UAV)-mounted cameras to predict binary pipe segmentation maps.

ICASSP_2024_Psarras_Poster.pdf

ICASSP_2024_Psarras_Poster.pdf (9)

Categories:: Other

9 Views

AUDIO-VISUAL SPEECH RECOGNITION IN-THE-WILD: MULTI-ANGLE VEHICLE CABIN CORPUS AND ATTENTION-BASED METHOD

Audio-Visual Speech Recognition In-The-Wild: Multi-Angle Vehicle Cabin Corpus And Attention-Based Method

ICASSP_2024_v3.pptx

ICASSP_2024_v3.pptx (15)

Categories:: Other

13 Views

Lightning Talk- Situation-Aware Tranmit Beamforming for Automotive radar

Read more about Lightning Talk- Situation-Aware Tranmit Beamforming for Automotive radar
Log in to post comments

Millimeter-wave radar is a common sensor modality used in automotive driving for target detection and perception. These radars can benefit from side information on the environment being sensed, such as lane topologies or data from other sensors. Existing radars do not leverage this information to adapt waveforms or perform prior-aware inference. In this paper, we model the side information as an occupancy map and design transmit beamformers that are customized to the map. Our method maximizes the probability of detection in regions with a higher uncertainty on the presence of a target.

Lightning Talk.pptx

Lightning Talk.pptx (11)

Categories:: Other

6 Views

Pages