1,720,968 research outputs found

    Video Processing and Communications Researches

    No full text
    video/mp4In this talk, my researches on video processing andcommunications conducted at video technology research group, Chulalongkorn University. There are several areas of researches. Firstly, wireless video coding and transmission researches arediscussed. The aim is to deliver very good quality video underchannel error constraint. Secondly, we have conducted many years ofresearches on video analytics for surveillance applications. Thirdly, current project that integrates the knowledge of video analytic, videocoding, and wireless communications is on-going. We are developed theprototype of video based security monitoring using high speed wirelesscommunication system. The fourth area is the accessible applicationswhich is Electronics Thai Sign Language Communication System. The last area is social media analysis on East Japan Earthquake which isthe collaboration with Gakushuin University, Japan.講演者所属: Chulalongkorn University講演日: 平成27年10月7日講演場所: 情報科学研究科大講義室L

    Multimodal Augmented Reality – Augmenting Auditory-Tactile Feedback to Change the Perception of Thickness

    Get PDF
    With vision being a primary sense of humans, we often first estimate the physical properties of objects by looking at them. However, when in doubt, for example, about the material they are made of or its structure, it is natural to apply other senses, such as haptics by touching them. Aiming at the ultimate goal of achieving a full-sensory augmented reality experience, we present an initial study focusing on multimodal feedback when tapping an object to estimate the thickness of its material. Our results indicate that we can change the perception of thickness of stiff objects by modulating acoustic stimuli. For flexible objects, which have a more distinctive tactile characteristic, adding vibratory responses when tapping on thick objects can make people perceive them as thin. We also identified that in the latter case, adding congruent acoustic stimuli does not further enhance the illusion but worsens it

    Joint FMO and Adaptive Intra Refresh for Error Resilience in H.264 Video Coding

    Get PDF
    In this paper, we propose an explicit Flexible Macroblock Ordering (FMO) map using distortion from error propagation. The effects caused by the damaged MBs of the current frame to the next frame and the other MBs in the same slice group are estimated. These effects can be used in the evaluation of the MBs' importance in the current frame, and a suitable map with a reduced effect of error propagation can be generated. In addition, to counteract with error propagation, intra refresh is a known method used to reduce the dependency between frames therefore could stop error propagation. Taking into account the channel state information, this paper also proposes an intra refresh algorithm that chooses the number of intra MBs for each frame. MBs having much effectiveness to the current frame and the next frame will be chosen for intra coding mode. By coupling this method with proposed FMO, it would help reduce the loss of important and intra coded MB. Results show that the proposed method has some improvements in terms of PSNR and the number of undecodable macroblocks as compared to some other methods.APSIPA ASC 2009: Asia-Pacific Signal and Information Processing Association, 2009 Annual Summit and Conference. 4-7 October 2009. Sapporo, Japan. Poster session: Image, Video, and Multimedia Signal Processing 2 (6 October 2009)

    Video Processing and Communications Researches

    No full text
    video/mp4In this talk, my researches on video processing andcommunications conducted at video technology research group, Chulalongkorn University. There are several areas of researches. Firstly, wireless video coding and transmission researches arediscussed. The aim is to deliver very good quality video underchannel error constraint. Secondly, we have conducted many years ofresearches on video analytics for surveillance applications. Thirdly, current project that integrates the knowledge of video analytic, videocoding, and wireless communications is on-going. We are developed theprototype of video based security monitoring using high speed wirelesscommunication system. The fourth area is the accessible applicationswhich is Electronics Thai Sign Language Communication System. The last area is social media analysis on East Japan Earthquake which isthe collaboration with Gakushuin University, Japan.講演者所属: Chulalongkorn University講演日: 平成27年10月7日講演場所: 情報科学研究科大講義室L1vide

    Reducing Complexity on Coding Unit Partitioning in Video Coding: A Review

    Get PDF
    In this article, we present a survey on the low complexity video coding on a coding unit (CU) partitioning with the aim for researchers to understand the foundation of video coding and fast CU partition algorithms. Firstly, we introduce video coding technologies by explaining the trending standards and reference models. They are High Efficiency Video Coding (HEVC), Joint Exploration Test Model (JEM), and VVC, which introduce novel quadtree (QT), quadtree plus binary tree (QTBT), quadtree plus multi-type tree (QTMT) block partitioning with expensive computation complexity, respectively. Secondly, we present a comprehensive explanation of the time-consuming CU partitioning, especially for researchers who are not familiar with CU partitioning. The newer the video coding standard, the more flexible partition structures and the higher the computational complexity. Then, we provide a deep and comprehensive survey of recent and state-of-the-art researches. Finally, we include a discussion section about the advantages and disadvantage of heuristic based and learning based approaches for the readers to explore quickly the performance of the existing algorithms and their limitations. To our knowledge, it is the first comprehensive survey to provide sufficient information about fast CU partitioning on HEVC, JEM, and VVC

    Going Beyond Counting First Authors in Author Co-citation Analysis

    Get PDF
    The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed

    An Advanced Features Extraction Module for Remote Sensing Image Super-Resolution

    Get PDF
    In recent years, convolutional neural networks (CNNs) have achieved remarkable advancement in the field of remote sensing image super-resolution due to the complexity and variability of textures and structures in remote sensing images (RSIs), which often repeat in the same images but differ across others. Current deep learning-based super-resolution models focus less on high-frequency features, which leads to suboptimal performance in capturing contours, textures, and spatial information. State-of-the-art CNN-based methods now focus on the feature extraction of RSIs using attention mechanisms. However, these methods are still incapable of effectively identifying and utilizing key content attention signals in RSIs. To solve this problem, we proposed an advanced feature extraction module called Channel and Spatial Attention Feature Extraction (CSA-FE) for effectively extracting the features by using the channel and spatial attention incorporated with the standard vision transformer (ViT). The proposed method trained over the UCMerced dataset on scales 2, 3, and 4. The experimental results show that our proposed method helps the model focus on the specific channels and spatial locations containing high-frequency information so that the model can focus on relevant features and suppress irrelevant ones, which enhances the quality of super-resolved images. Our model achieved superior performance compared to various existing models.Comment: Preprint of paper from The 21st International Conference on Electrical Engineering/Electronics, Computer, Telecommunications and Information Technology or ECTI-CON 2024, Khon Kaen, Thailan
    corecore