Search Results for author: Masahiro Yasuda

Found 13 papers, 6 papers with code

Guided Masked Self-Distillation Modeling for Distributed Multimedia Sensor Event Analysis

no code implementations12 Apr 2024 Masahiro Yasuda, Noboru Harada, Yasunori Ohishi, Shoichiro Saito, Akira Nakayama, Nobutaka Ono

This is because the information obtained from a single sensor is often missing or fragmented in such an environment; observations from multiple locations and modalities should be integrated to analyze events comprehensively.

6DoF SELD: Sound Event Localization and Detection Using Microphones and Motion Tracking Sensors on self-motioning human

no code implementations4 Mar 2024 Masahiro Yasuda, Shoichiro Saito, Akira Nakayama, Noboru Harada

A system trained only with a dataset using microphone arrays in a fixed position would be unable to adapt to the fast relative motion of sound events associated with self-motion, resulting in the degradation of SELD performance.

Sound Event Localization and Detection

First-shot anomaly sound detection for machine condition monitoring: A domain generalization baseline

1 code implementation1 Mar 2023 Noboru Harada, Daisuke Niizumi, Yasunori Ohishi, Daiki Takeuchi, Masahiro Yasuda

This paper provides a baseline system for First-shot-compliant unsupervised anomaly detection (ASD) for machine condition monitoring.

Domain Generalization Task 2 +1

Multi-view and Multi-modal Event Detection Utilizing Transformer-based Multi-sensor fusion

1 code implementation18 Feb 2022 Masahiro Yasuda, Yasunori Ohishi, Shoichiro Saito, Noboru Harada

We tackle a challenging task: multi-view and multi-modal event detection that detects events in a wide-range real environment by utilizing data from distributed cameras and microphones and their weak labels.

Event Detection Sensor Fusion

APPLADE: Adjustable Plug-and-play Audio Declipper Combining DNN with Sparse Optimization

no code implementations16 Feb 2022 Tomoro Tanaka, Kohei Yatabe, Masahiro Yasuda, Yasuhiro Oikawa

Still, they cannot perform well if the training data have mismatches and/or constraints in the time domain are not imposed.

Audio declipping

A Transformer-based Audio Captioning Model with Keyword Estimation

no code implementations1 Jul 2020 Yuma Koizumi, Ryo Masumura, Kyosuke Nishida, Masahiro Yasuda, Shoichiro Saito

TRACKE estimates keywords, which comprise a word set corresponding to audio events/scenes in the input audio, and generates the caption while referring to the estimated keywords to reduce word-selection indeterminacy.

Acoustic Scene Classification Audio captioning +2

DOA Estimation by DNN-based Denoising and Dereverberation from Sound Intensity Vector

no code implementations10 Oct 2019 Masahiro Yasuda, Yuma Koizumi, Luca Mazzon, Shoichiro Saito, Hisashi Uematsu

We propose a direction of arrival (DOA) estimation method that combines sound-intensity vector (IV)-based DOA estimation and DNN-based denoising and dereverberation.

Denoising

Cannot find the paper you are looking for? You can Submit a new open access paper.