arXiv cs.AI / cs.LG / cs.CL·2d agoTrackEverything: Long Horizon Dense Tracking via De-Duplicating 3D Scene Representations#3d-tracking#computer-vision#point-trackingAI research
Hugging Face daily papers·3d agoRGBD20K: A Large-Scale Benchmark for RGB-D Semantic Segmentation#rgb-d#semantic-segmentation#benchmark
arXiv cs.CR·6d agoReinforcement Learning Inspired Black-box Adversarial Attacks for Computer Vision#adversarial-attack#black-box#reinforcement-learningAI safety & security
Hugging Face daily papers·10d agoTraining-Adaptive Convolutional Sparse Coding via Information Bottleneck for Robust Visual Representation#computer-vision#information-bottleneck#representation-learning
Hugging Face daily papers·10d agoRefinement Is Inherently Editable: Training-Free Prompt-to-Prompt Image Editing with Generative Refinement Network#computer-vision#diffusion-models#image-editing
Meta Newsroom·10d agoCanadian Start-up smartARM Uses AI to Create Intuitive Bionic Prosthetics#ai-glasses#computer-vision#dinov2 2 min
Hugging Face daily papers·12d agoTAPe+ML: A Compact Structured Representation for Multi-Task Computer Vision#compact-models#computer-vision#instance-segmentation
Hugging Face daily papers·12d agoEventEgoHands++: Event-based Egocentric 3D Hand Mesh Reconstruction with Real Dataset#ar-vr#computer-vision#dataset
Hugging Face daily papers·16d agoRelateAnything: Real-Time Open-Vocabulary Relation Prediction From Any Inputs#benchmark#computer-vision#dataset1
Hugging Face daily papers·17d agoFeature Recovery for Object Understanding After Irreversible Fire Damage#benchmark#computer-vision#degradation
Hugging Face daily papers·17d agoFreeFlow: A Bias-free Hierarchical Transformer for Optical Flow Estimation#benchmarks#computer-vision#deep-learning1
Hugging Face daily papers·19d agoSynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation#benchmark#computer-vision#dataset
Hugging Face daily papers·19d agoMarigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation#computer-vision#depth-estimation#diffusion-transformers
Hugging Face daily papers·20d agoRelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting#computer-vision#generative-models#multi-view
Hugging Face daily papers·21d agoOracleZoom: On-Policy Self-Distillation Inspired Reference-Constrained Recursive Image Super Resolution#computer-vision#hallucination#image-quality
Hugging Face daily papers·24d agoAdaptVPR: Route-Aware Hard Positive Generation for Robust Visual Place Recognition#adaptcities#computer-vision#data-augmentation
Help Net Security·25d agoAn AI CAPTCHA solver talked itself out of the right answer#benchmark#captcha#computer-vision 3 min1