Everything you care about in one place

Follow feeds: blogs, news, RSS and more. An effortless way to read and digest content of your choice.

Get Feeder

arxiv.org

cs.CV updates on arXiv.org

Get the latest updates from cs.CV updates on arXiv.org directly as they happen.

Follow now 114 followers

Latest posts

Last updated about 22 hours ago

Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models

about 22 hours ago

arXiv:2605.16409v3 Announce Type: replace Abstract: Optical character recognition (OCR) and multilingual scene-text...

CalibAnyView: Beyond Single-View Camera Calibration in the Wild

about 22 hours ago

arXiv:2605.14615v2 Announce Type: replace Abstract: Camera calibration is fundamental to reliable geometric...

Hyper-FSAD: Training-Free and Language-Free Few-Shot Anomaly Detection via Sparse Hyper Matching

about 22 hours ago

arXiv:2605.10628v2 Announce Type: replace Abstract: Few-shot anomaly detection (FSAD) is particularly valuable...

Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization

about 22 hours ago

arXiv:2605.09996v2 Announce Type: replace Abstract: While multimodal large language models have advanced...

simpleposter: A simple baseline for product poster generation

about 22 hours ago

arXiv:2605.08784v2 Announce Type: replace Abstract: Product poster generation poses distinct challenges beyond...

Significance and Stability Analysis of Genotype-Environment Interaction using GxEStat

about 22 hours ago

arXiv:2604.03337v3 Announce Type: replace Abstract: Genotype-environment (GxE) interactions can influence the performance...

A Simple Efficiency Incremental Learning Framework via Vision-Language Model with Nonlinear Multi-Adapters

about 22 hours ago

arXiv:2603.11211v3 Announce Type: replace Abstract: Incremental Learning (IL) aims to learn new...

TC-Pad\'e: Trajectory-Consistent Pad\'e Approximation for Diffusion Acceleration

about 22 hours ago

arXiv:2603.02943v2 Announce Type: replace Abstract: Despite achieving state-of-the-art generation quality, diffusion models...

Diffusion Probe: Generated Image Result Prediction Using CNN Probes

about 22 hours ago

arXiv:2602.23783v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models lack an efficient...

ForgeryVCR: Visual-Centric Reasoning via Efficient Forensic Tools in MLLMs for Image Forgery Detection and Localization

about 22 hours ago

arXiv:2602.14098v2 Announce Type: replace Abstract: Existing Multimodal Large Language Models (MLLMs) for...

SCAR-GS: Spatial Context Attention for Residuals in Progressive Gaussian Splatting

about 22 hours ago

arXiv:2601.04348v2 Announce Type: replace Abstract: Recent advances in 3D Gaussian Splatting have...

Self-Supervised Weighted Image Guided Quantitative MRI Super-Resolution

about 22 hours ago

arXiv:2512.17612v2 Announce Type: replace Abstract: Object: To present and evaluate Self-supervised Weighted...