Description Language Research Articles

The study of practice based on visualization and the role of visualization in shaping the outcomes of contemplative practice is an overlooked research niche. The aim of this article is to contribute to a deeper understanding of the dynamics of the process that connects literary images and visual perception in relevant meditative practices. The key question is which elements contribute to forming this connection in contemplative practices like meditation. To gain a deeper understanding of the dynamics involved, the study employed a mix of primary and secondary data sources, along with analytical, descriptive, and phenomenological methods, drawing from an interdisciplinary approach that includes psychology, neuroscience, and literature. The following meditation practices were selected: contemplative meditation, creative visualization, koan meditation, imaginative meditation, creative workshops with haiku poetry as an outcome, and meditative storytelling. The identified elements in the connection between literary depictions and visual perception include mental imagery and visualization, cognitive processes, sensory processes, and emotional processes. The importance of literary techniques is particularly highlighted – using descriptive language with metaphors, similes, allegories, and other stylistic figures that create strong visual images. This connection has a neurological basis – neuroscience studies show that reading/listening to texts describing visual experiences activates brain areas involved in actual visual perception. This overlap suggests that the brain processes literary descriptions similarly to how it processes real visual stimuli. The connection between literary depictions and visual perception significantly enhances the quality of meditative practice and promotes deeper understanding and emotional-volitional engagement with the text and personal development of the meditators. Both secondary and primary data sources were used in the paper.

Read full abstract

Pattern recognition through the fusion of RGB frames and Event streams has emerged as a novel research area in recent years. Current methods typically employ backbone networks to individually extract the features of RGB frames and event streams, and subsequently fuse these features for pattern recognition. However, we posit that these methods may suffer from two key issues: (1). They attempt to directly learn a mapping from the input vision modality to the semantic labels. This approach often leads to sub-optimal results due to the disparity between the input and semantic labels; (2). They utilize small-scale backbone networks for the extraction of RGB and Event input features, thus these models fail to harness the recent performance advancements of large-scale visual-language models. In this study, we introduce a novel pattern recognition framework that consolidates the semantic labels, RGB frames, and event streams, leveraging pre-trained large-scale vision–language models. Specifically, given the input RGB frames, event streams, and all the predefined semantic labels, we employ a pre-trained large-scale vision model (CLIP vision encoder) to extract the RGB and event features. To handle the semantic labels, we initially convert them into language descriptions through prompt engineering and polish using ChatGPT, and then obtain the semantic features using the pre-trained large-scale language model (CLIP text encoder). Subsequently, we integrate the RGB/Event features and semantic features using multimodal Transformer networks. The resulting frame and event tokens are further amplified using self-attention layers. Concurrently, we propose to enhance the interactions between text tokens and RGB/Event tokens via cross-attention. Finally, we consolidate all three modalities using self-attention and feed-forward layers for recognition. Comprehensive experiments on the HARDVS and PokerEvent datasets fully substantiate the efficacy of our proposed SAFE model. The source code has been released at https://github.com/Event-AHU/SAFE_LargeVLM.

Read full abstract

Description Language Research Articles

Related Topics

Articles published on Description Language

Image Captioning Based on Semantic Scenes.

The Pedagogical Tipiṭaka: OER & the Three Baskets of Ancient Language Instruction

Hardware-accelerated neural network model for early prediction of sudden cardiac arrest based on heart rate variability metrics

The Connection Between Literary Images and Visual Perception in Meditative Practice

Text-and-Image Learning Transformer for Cross-modal Person Re-identification

Increasing the speed of digital devices using the ASMD-FSMD method using non-blocking operators

Semantic-aware frame-event fusion based pattern recognition via large vision–language models

UART IP core Design and Verification using WISHBONE Interface

MapGPT: an autonomous framework for mapping by integrating large language model and cartographic tools

Algorithm-Driven Robotic Discovery of Polyoxometalate-Scaffolding Metal-Organic Frameworks.

Automating Comment Generation for Smart Contract from Bytecode

Flight Arrival Scheduling via Large Language Model

An early parenting intervention focused on enriched parent-child interactions improves effortful control in the early years of school.

Advancing Zebrafish Husbandry: Takeaways From the 2024 Husbandry Workshop and Husbandry Summit.

Interactive Surgical Training in Neuroendoscopy: Real-Time Anatomical Feature Localization Using Natural Language Expressions.

Extraction of logical networks during decompiling transistor-level CMOS circuit descriptions

A Transfer Learning Approach for Arabic Image Captions

An efficient XOR-free implementation of polar encoder for reconfigurable hardware

Referring Image Segmentation with Multi-Modal Feature Interaction and Alignment Based on Convolutional Nonlinear Spiking Neural Membrane Systems.

Linguacodus: a synergistic framework for transformative code generation in machine learning pipelines

Lead the way for us

Editage

Paperpal

R Discovery

Mind the Graph

Description Language Research Articles

Related Topics

Articles published on Description Language

Image Captioning Based on Semantic Scenes.

The Pedagogical Tipiṭaka: OER &amp; the Three Baskets of Ancient Language Instruction

Hardware-accelerated neural network model for early prediction of sudden cardiac arrest based on heart rate variability metrics

The Connection Between Literary Images and Visual Perception in Meditative Practice

Text-and-Image Learning Transformer for Cross-modal Person Re-identification

Increasing the speed of digital devices using the ASMD-FSMD method using non-blocking operators

Semantic-aware frame-event fusion based pattern recognition via large vision–language models

UART IP core Design and Verification using WISHBONE Interface

MapGPT: an autonomous framework for mapping by integrating large language model and cartographic tools

Algorithm-Driven Robotic Discovery of Polyoxometalate-Scaffolding Metal-Organic Frameworks.

Automating Comment Generation for Smart Contract from Bytecode

Flight Arrival Scheduling via Large Language Model

An early parenting intervention focused on enriched parent-child interactions improves effortful control in the early years of school.

Advancing Zebrafish Husbandry: Takeaways From the 2024 Husbandry Workshop and Husbandry Summit.

Interactive Surgical Training in Neuroendoscopy: Real-Time Anatomical Feature Localization Using Natural Language Expressions.

Extraction of logical networks during decompiling transistor-level CMOS circuit descriptions

A Transfer Learning Approach for Arabic Image Captions

An efficient XOR-free implementation of polar encoder for reconfigurable hardware

Referring Image Segmentation with Multi-Modal Feature Interaction and Alignment Based on Convolutional Nonlinear Spiking Neural Membrane Systems.

Linguacodus: a synergistic framework for transformative code generation in machine learning pipelines

The Pedagogical Tipiṭaka: OER & the Three Baskets of Ancient Language Instruction