Yohaku Film Research Papers
Working Paper Series / YFRI-WP-2026-002Attention Convergence Model
A Hypothesis on the Cognitive Origin of Perception
Abstract
This paper proposes the Attention Convergence Model (ACM), a theoretical hypothesis describing the earliest stage of human perception.
Previous work introduced the Reversible State Design Model, which proposed that creators and viewers are connected through a shared cognitive State.
The present study investigates a stage preceding State formation and hypothesizes that perception begins not with object recognition, meaning, or narrative, but with the convergence of attention.
A minimal visual stimulus may narrow the viewer’s attention toward a single point, temporarily suppress external distractions and allowing a quiet cognitive State to emerge.
This paper proposes that Attention precedes State, providing a possible foundation for future studies in perception, visual communication, State Design, and AI-assisted expression.
Introduction
Many visual communication models implicitly assume that perception begins when a viewer recognizes an object, scene, or message.
However, observation during visual design experiments suggests that a more fundamental cognitive event may occur before recognition.
When presented with a minimal and ambiguous stimulus, such as a narrow beam of light within an otherwise quiet visual field, the viewer may first experience a concentration of attention.
Perception may begin before the viewer knows what is being perceived.
This study examines the hypothesis that attention convergence is the cognitive entrance through which State, perception, meaning, and narrative subsequently emerge.
Conventional Perception Model
A simplified conventional model may be represented as follows:
Let X represent a visual stimulus, O an identified object, and M the meaning attributed to that object.
While useful for describing established recognition, this sequence does not fully describe what occurs when a viewer initially encounters a stimulus that has not yet been classified.
Proposed Attention Convergence Model
The proposed model introduces Attention and State before object recognition and meaning formation.
Here:
- X represents the initial stimulus;
- A represents attention convergence;
- S represents the resulting cognitive State;
- P represents perception and recognition;
- M represents generated meaning;
- N represents the emerging narrative.
Perception does not begin with recognition.
Perception begins when attention converges.
Attention Convergence
Attention convergence is defined here as the temporary narrowing of cognitive focus toward a limited visual stimulus.
For example, a viewer encountering a narrow beam of light may not immediately interpret the stimulus as light.
Instead, the viewer may first experience a non-verbal cognitive reaction:
What is this?
The direction, source, boundary, and nature of the stimulus remain uncertain. This uncertainty may cause the viewer’s attention to concentrate more strongly.
In this conceptual notation:
- ΔI represents perceptual intensity difference;
- U represents uncertainty or ambiguity;
- C represents contrast;
- D represents competing external distraction.
This equation is conceptual and does not yet constitute an empirically validated measurement function.
Cognitive Silence
A quiet cognitive State may emerge when attention becomes sufficiently concentrated.
The stimulus does not directly produce silence. Instead, silence may be experienced because competing external awareness is temporarily reduced.
Here, Sq represents a quiet cognitive State, A represents focused attention, D represents distraction, and ε is a conceptual stability term.
Perception Sequence
Preliminary observation suggests that the Viewer may construct a perceptual world through the following sequence.
This sequence describes the gradual internal construction of a perceptual world rather than a direct reception of completed meaning.
Inverse Creator Sequence
The Creator may proceed in the opposite direction by decomposing an internal narrative into increasingly minimal perceptual elements.
The Creator begins with an internal story or intention, determines the desired meaning and cognitive State, selects an existence or object, defines its boundaries, and reduces the final expression to a stimulus capable of gathering attention.
The Viewer constructs a world.
The Creator decomposes one.
Attention as the Entrance to State Design
Previous work proposed that creative expression should target a cognitive State rather than a predetermined emotional result.
The present model extends that hypothesis by proposing that State cannot be established without first directing attention.
Before a State can be designed, attention must first be gathered.
Implications for Artificial Intelligence
Current generative AI systems primarily optimize visible output, semantic relevance, or aesthetic similarity.
The Attention Convergence Model suggests an additional design layer: the intended distribution and concentration of Viewer attention.
AI would therefore function not merely as an image generator but as an Attention and State Expression Transformer.
Viewer behavior and cognitive feedback could subsequently be used to refine expression parameters.
This expression is conceptual and does not yet define a validated optimization algorithm.
Future Research
Future research should investigate:
- whether attention convergence consistently precedes object recognition;
- how visual ambiguity affects attention duration and intensity;
- the relationship between focused attention and cognitive silence;
- whether minimal visual stimuli suppress awareness of external distractions;
- how boundary recognition develops after attention convergence;
- methods for measuring Viewer attention without disrupting the experience;
- the relationship between Attention Design and State Design;
- AI-assisted optimization of attention and State parameters;
- applications to film, Web design, photography, architecture, UI, branding, exhibition, and spatial experience.
Discussion
The present hypothesis suggests that the first role of visual expression may not be communication.
Its first role may be to produce a perceptual difference capable of gathering attention.
Only after attention converges can a viewer enter a State, detect a boundary, recognize an existence, generate meaning, and construct a narrative.
This implies that minimal expression is not necessarily the absence of information. It may function as an active condition through which cognitive focus becomes possible.
The Viewer may not initially observe the work.
The Viewer may first experience attention itself.
Conclusion
This paper proposes the Attention Convergence Model as a hypothesis concerning the cognitive origin of perception.
The model suggests that perception begins not with object recognition, but with the convergence of attention toward a perceptual difference.
Attention may temporarily reduce competing external awareness, allowing a quiet cognitive State to emerge.
From this State, boundaries, existence, meaning, and narrative may be progressively constructed.
Attention is not merely directed toward perception.
Attention may be the condition from which perception begins.
Although the model remains conceptual, it extends the Reversible State Design Model toward the earliest stage of Viewer cognition and provides a possible foundation for Attention Design, State Design, and AI-assisted expression research.
Relationship to Previous Work
YFRI Working Paper Series
Paper 01 — Reversible State Design Model
Proposed that Creator and Viewer are connected through State.
Paper 02 — Attention Convergence Model
Proposes that Attention precedes State and forms the cognitive
entrance to perception.
Together, the papers form an emerging theoretical framework linking expression, attention, State, perception, and meaning generation.