Spatial-Analysis is a curated collection of vision-language models for spatial reasoning across multiple images. It includes models designed to answer
sabarinathan
sabaridsnfuji
AI & ML interests
Hi, I’m Sabarinathan, an AI Engineer and developer passionate about building robust and responsible multimodal models, including OCR and document understanding for English, Japanese, and Tamil.
My work spans speech recognition, image understanding, and transformer-based architectures.
Interests:
Vision-Language Models (VLM) and lightweight computer vision models
Natural Language Processing (NLP)
Automatic Speech Recognition (ASR)
Recent Activity
updated a model 14 days ago
sabaridsnfuji/Qwen3-VL-4B-COMPASS-T-DINOv3 published a model 14 days ago
sabaridsnfuji/Qwen3-VL-4B-COMPASS-T-DINOv3 commentedon an article 21 days ago
What We Learned by Reproducing 2,200 papers from ICML