首页 > 专栏 > cs.CV updates on arXiv.org cs.CV updates on arXiv.org 共 5006 条资讯 DSSR-3D: Decoupled Reasoning for View-Dependent Referring in 3D Gaussians 2026-06-28 03:08:04 Seeing the City or Recognizing the Place? What Street-View Imagery Adds Beyond Existing Urban Data in VLM Urban Sensing 2026-06-28 03:08:04 Hardware-Algorithm Co-Optimization of Early-Exit Neural Networks for Multi-Core Edge Accelerators 2026-06-28 03:08:04 Reachability Is Not Generalization: Understanding Verb--Noun Decomposition in Assembly Action Recognition 2026-06-28 03:08:04 A Comprehensive Review of One-Pixel Attack: Research Status, Taxonomy, Applications, Regulation Policy and Future Directions 2026-06-28 03:08:04 MIRTO: a registration-gated, multiverse-tested evaluation protocol for unsupervised anomaly segmentation in brain MRI 2026-06-28 03:08:04 CoEvolve: Construct-to-Edit Visual Grounding with Bidirectional State Refinement 2026-06-28 03:08:04 Spatially Gated Diffusion for Localized Counterfactual Chest Radiograph Editing 2026-06-28 03:08:04 Augmented Equivariant Mesh Networks for Anatomical Segmentation 2026-06-28 03:08:04 DisasterInsight: A Building-Centric Benchmark for Evaluating Vision--Language Models in Disaster Response 2026-06-28 03:08:04 PhysVista: Benchmarking Physical Intelligence in VLMs via a Perception-Reasoning-Assessment Loop 2026-06-28 03:08:04 Memorizon: Training World Models Beyond Their Context Window 2026-06-28 03:08:04 DexPolicy: Scheduled Exploration for Trajectory-Guided Dexterous Manipulation 2026-06-28 03:08:04 PixelDense: Dense Prediction as Representation Alignment for Pixel Diffusion 2026-06-28 03:08:04 Align Then Reason: A Multimodal Lip-Sync Judge for Dubbing 2026-06-28 03:08:04 Personalized Image Generation with Reasoning and Reflection 2026-06-28 03:08:04 OmniSeek: Native Tool Integration for Multi-turn Audio-Visual Reasoning 2026-06-28 03:08:04 World Observer: Joint Actor-Observer Generation for Persistent World Modeling 2026-06-28 03:08:04 Distill the Visual Evidence, Not Just the Answer: Cross-World On-Policy Distillation for Vision-Language Models 2026-06-28 03:08:04 Correcting WHERE, Preserving HOW: Compositional Generalization for Vision-Language-Action Models via Referential Guidance 2026-06-28 03:08:04 « 上一页1…34567…251下一页 » 相关分类 #!/slash/note #UNTAG (B)(F)uzzing on my world (Hi)story (IN)SECURE Magazine Notification (gdb) break *0x972 - 带鱼博客 BeltfishBlog - ./kwaa.dev .NET Blog .Trash /home/rook1e 00's Adventure 0kami's Blog 0x41414141 in ?? () 0x7f Blog 0xRick Owned Root ! 0xd00's blog 1 Byte 1A23 Blog 1A23 Studio 1Link.Fun 1stwebdesigner 251 2BAB 的工程博客 2ch中文网 360 CERT 360 Netlab Blog - Network Securi 38号车评中心 3o米的微博 404 Media