In this paper, we propose a novel methodology aimed at simulating the learning phenomenon of nystagmus through the application of differential blurring on datasets. Nystagmus is a biological phenomenon that influences human vision throughout life, notably by diminishing head shake from infancy to adulthood. Leveraging this concept, we address the issue of waste classification, a pressing global concern. The proposed framework comprises two modules, with the second module closely resembling the original Vision transformer, a state-of-the-art model model in classification tasks. The primary motivation behind our approach is to enhance the precision and adaptability of intelligent systems, mirroring the real-world conditions that the human visual system undergoes. This novel methodology surpasses the standard Vision Transformer model in waste classification tasks, exhibiting an improvement with a margin of 2–6%. This improvement underscores the potential of our methodology in improving the precision of artificially intelligent models by drawing inspiration from human vision perception. Further research in the proposed methodology could yield greater performance results, and can be extrapolated to other global issues and intelligent systems.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Integrating Human Vision Perception in Vision Transformers for Classifying Waste Items

  • Akshat Kishore Shrivastava,
  • Tapan Kumar Gandhi

摘要

In this paper, we propose a novel methodology aimed at simulating the learning phenomenon of nystagmus through the application of differential blurring on datasets. Nystagmus is a biological phenomenon that influences human vision throughout life, notably by diminishing head shake from infancy to adulthood. Leveraging this concept, we address the issue of waste classification, a pressing global concern. The proposed framework comprises two modules, with the second module closely resembling the original Vision transformer, a state-of-the-art model model in classification tasks. The primary motivation behind our approach is to enhance the precision and adaptability of intelligent systems, mirroring the real-world conditions that the human visual system undergoes. This novel methodology surpasses the standard Vision Transformer model in waste classification tasks, exhibiting an improvement with a margin of 2–6%. This improvement underscores the potential of our methodology in improving the precision of artificially intelligent models by drawing inspiration from human vision perception. Further research in the proposed methodology could yield greater performance results, and can be extrapolated to other global issues and intelligent systems.