Computer Vision and Pattern Recognition
Authors and titles for recent submissions
[ total of 763 entries: 1-104 | 105-208 | 209-312 | 313-416 | ... | 729-763 ][ showing 104 entries per page: fewer | more | all ]
Thu, 28 Mar 2024 (showing first 104 of 136 entries)
- [1] arXiv:2403.18820 [pdf, other]
-
Title: MetaCap: Meta-learning Priors from Multi-View Imagery for Sparse-view Human Performance Capture and RenderingComments: Project page: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [2] arXiv:2403.18819 [pdf, other]
-
Title: Benchmarking Object Detectors with COCO: A New Path ForwardSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [3] arXiv:2403.18818 [pdf, other]
-
Title: ObjectDrop: Bootstrapping Counterfactuals for Photorealistic Object Removal and InsertionSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [4] arXiv:2403.18816 [pdf, other]
-
Title: Garment3DGen: 3D Garment Stylization and Texture GenerationComments: Project Page: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [5] arXiv:2403.18814 [pdf, other]
-
Title: Mini-Gemini: Mining the Potential of Multi-modality Vision Language ModelsAuthors: Yanwei Li, Yuechen Zhang, Chengyao Wang, Zhisheng Zhong, Yixin Chen, Ruihang Chu, Shaoteng Liu, Jiaya JiaComments: Code and models are available at this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
- [6] arXiv:2403.18811 [pdf, other]
-
Title: Duolando: Follower GPT with Off-Policy Reinforcement Learning for Dance AccompanimentAuthors: Li Siyao, Tianpei Gu, Zhitao Yang, Zhengyu Lin, Ziwei Liu, Henghui Ding, Lei Yang, Chen Change LoyComments: ICLR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Graphics (cs.GR); Sound (cs.SD); Audio and Speech Processing (eess.AS)
- [7] arXiv:2403.18807 [pdf, other]
-
Title: ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth EstimationComments: Accepted at IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [8] arXiv:2403.18795 [pdf, other]
-
Title: Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstructionSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [9] arXiv:2403.18791 [pdf, other]
-
Title: Object Pose Estimation via the Aggregation of Diffusion FeaturesComments: Accepted to CVPR2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [10] arXiv:2403.18784 [pdf, other]
-
Title: SplatFace: Gaussian Splat Face Reconstruction Leveraging an Optimizable SurfaceSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [11] arXiv:2403.18775 [pdf, other]
-
Title: ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic ObjectComments: Accepted at CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [12] arXiv:2403.18774 [pdf, other]
-
Title: RAW: A Robust and Agile Plug-and-Play Watermark Framework for AI-Generated Images with Provable GuaranteesSubjects: Computer Vision and Pattern Recognition (cs.CV); Cryptography and Security (cs.CR); Machine Learning (cs.LG)
- [13] arXiv:2403.18762 [pdf, other]
-
Title: ModaLink: Unifying Modalities for Efficient Image-to-PointCloud Place RecognitionAuthors: Weidong Xie, Lun Luo, Nanfei Ye, Yi Ren, Shaoyi Du, Minhang Wang, Jintao Xu, Rui Ai, Weihao Gu, Xieyuanli ChenComments: 8 pages, 11 figures, conferenceSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Robotics (cs.RO)
- [14] arXiv:2403.18756 [pdf, ps, other]
-
Title: Detection of subclinical atherosclerosis by image-based deep learning on chest x-rayAuthors: Guglielmo Gallone, Francesco Iodice, Alberto Presta, Davide Tore, Ovidio de Filippo, Michele Visciano, Carlo Alberto Barbano, Alessandro Serafini, Paola Gorrini, Alessandro Bruno, Walter Grosso Marra, James Hughes, Mario Iannaccone, Paolo Fonio, Attilio Fiandrotti, Alessandro Depaoli, Marco Grangetto, Gaetano Maria de Ferrari, Fabrizio D'AscenzoComments: Submitted to European Heart Journal - Cardiovascular Imaging Added also the additional material 44 pages (30 main paper, 14 additional material), 14 figures (5 main manuscript, 9 additional material)Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [15] arXiv:2403.18730 [pdf, other]
-
Title: Towards Image Ambient Lighting NormalizationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [16] arXiv:2403.18715 [pdf, other]
-
Title: Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive DecodingSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Multimedia (cs.MM)
- [17] arXiv:2403.18714 [pdf, other]
- [18] arXiv:2403.18711 [pdf, other]
-
Title: SAT-NGP : Unleashing Neural Graphics Primitives for Fast Relightable Transient-Free 3D reconstruction from Satellite ImageryComments: 5 pages, 3 figures, 1 table; Accepted to International Geoscience and Remote Sensing Symposium (IGARSS) 2024; Code available at this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [19] arXiv:2403.18708 [pdf, other]
-
Title: Dense Vision Transformer Compression with Few SamplesComments: Accepted to CVPR 2024. Note: Jianxin Wu is a contributing author for the arXiv version of this paper but is not listed as an author in the CVPR version due to his role as Program ChairSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [20] arXiv:2403.18690 [pdf, other]
-
Title: Annolid: Annotate, Segment, and Track Anything You NeedSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [21] arXiv:2403.18674 [pdf, other]
-
Title: Deep Learning for Robust and Explainable Models in Computer VisionAuthors: Mohammadreza AmirianComments: 150 pages, 37 figures, 12 tablesJournal-ref: OPARU is the OPen Access Repository of Ulm University and Ulm University of Applied Sciences, 2023Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [22] arXiv:2403.18649 [pdf, other]
-
Title: Addressing Data Annotation Challenges in Multiple Sensors: A Solution for Scania Collected DatasetsAuthors: Ajinkya Khoche, Aron Asefaw, Alejandro Gonzalez, Bogdan Timus, Sina Sharif Mansouri, Patric JensfeltComments: Accepted to European Control Conference 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Systems and Control (eess.SY)
- [23] arXiv:2403.18605 [pdf, other]
-
Title: FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image EditingComments: Our project page: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [24] arXiv:2403.18600 [pdf, other]
-
Title: RAP: Retrieval-Augmented Planner for Adaptive Procedure Planning in Instructional VideosComments: 23 pages, 6 figures, 12 tablesSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Robotics (cs.RO)
- [25] arXiv:2403.18593 [pdf, other]
-
Title: Homogeneous Tokenizer Matters: Homogeneous Visual Tokenizer for Remote Sensing Image UnderstandingComments: 20 pages, 8 figures, 6 tablesSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [26] arXiv:2403.18575 [pdf, other]
-
Title: HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object InteractionsSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [27] arXiv:2403.18565 [pdf, other]
-
Title: Artifact Reduction in 3D and 4D Cone-beam Computed Tomography Images with Deep Learning -- A ReviewComments: 16 pages, 4 figures, 1 Table, published in IEEE Access JournalJournal-ref: IEEE Access, vol. 12, pp. 10281-10295, 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [28] arXiv:2403.18554 [pdf, other]
-
Title: CosalPure: Learning Concept from Group Images for Robust Co-Saliency DetectionComments: 8 pagesSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [29] arXiv:2403.18551 [pdf, other]
-
Title: Attention Calibration for Disentangled Text-to-Image PersonalizationComments: Accepted to CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [30] arXiv:2403.18550 [pdf, other]
-
Title: OrCo: Towards Better Generalization via Orthogonality and Contrast for Few-Shot Class-Incremental LearningSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [31] arXiv:2403.18548 [pdf, other]
-
Title: A Semi-supervised Nighttime Dehazing Baseline with Spatial-Frequency Aware and Realistic Brightness ConstraintComments: This paper is accepted by CVPR2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [32] arXiv:2403.18525 [pdf, other]
-
Title: Language Plays a Pivotal Role in the Object-Attribute Compositional Generalization of CLIPComments: Oral accepted at OODCV 2023(this http URL)Subjects: Computer Vision and Pattern Recognition (cs.CV); Computation and Language (cs.CL); Machine Learning (cs.LG)
- [33] arXiv:2403.18512 [pdf, other]
-
Title: ParCo: Part-Coordinating Text-to-Motion SynthesisSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [34] arXiv:2403.18495 [pdf, other]
-
Title: Direct mineral content prediction from drill core images via transfer learningAuthors: Romana Boiger, Sergey V. Churakov, Ignacio Ballester Llagaria, Georg Kosakowski, Raphael Wüst, Nikolaos I. PrasianakisSubjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Image and Video Processing (eess.IV)
- [35] arXiv:2403.18493 [pdf, other]
-
Title: VersaT2I: Improving Text-to-Image Models with Versatile RewardAuthors: Jianshu Guo, Wenhao Chai, Jie Deng, Hsiang-Wei Huang, Tian Ye, Yichen Xu, Jiawei Zhang, Jenq-Neng Hwang, Gaoang WangSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [36] arXiv:2403.18490 [pdf, other]
-
Title: I2CKD : Intra- and Inter-Class Knowledge Distillation for Semantic SegmentationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [37] arXiv:2403.18476 [pdf, other]
-
Title: Modeling uncertainty for Gaussian SplattingSubjects: Computer Vision and Pattern Recognition (cs.CV); Graphics (cs.GR)
- [38] arXiv:2403.18471 [pdf, other]
-
Title: DiffusionFace: Towards a Comprehensive Dataset for Diffusion-Based Face Forgery AnalysisSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [39] arXiv:2403.18469 [pdf, other]
-
Title: Density-guided Translator Boosts Synthetic-to-Real Unsupervised Domain Adaptive Segmentation of 3D Point CloudsComments: CVPR2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [40] arXiv:2403.18461 [pdf, other]
-
Title: DiffStyler: Diffusion-based Localized Image Style TransferAuthors: Shaoxu LiSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [41] arXiv:2403.18454 [pdf, other]
-
Title: Scaling Vision-and-Language Navigation With Offline RLComments: Published in Transactions on Machine Learning Research (04/2024)Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [42] arXiv:2403.18452 [pdf, other]
-
Title: SingularTrajectory: Universal Trajectory Predictor Using Diffusion ModelComments: Accepted at CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Robotics (cs.RO)
- [43] arXiv:2403.18443 [pdf, other]
-
Title: $\mathrm{F^2Depth}$: Self-supervised Indoor Monocular Depth Estimation via Optical Flow Consistency and Feature Map SynthesisSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [44] arXiv:2403.18442 [pdf, other]
-
Title: Backpropagation-free Network for 3D Test-time AdaptationAuthors: Yanshuo Wang, Ali Cheraghian, Zeeshan Hayder, Jie Hong, Sameera Ramasinghe, Shafin Rahman, David Ahmedt-Aristizabal, Xuesong Li, Lars Petersson, Mehrtash HarandiComments: CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [45] arXiv:2403.18425 [pdf, other]
-
Title: U-Sketch: An Efficient Approach for Sketch to Image Diffusion ModelsSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [46] arXiv:2403.18417 [pdf, other]
-
Title: ECNet: Effective Controllable Text-to-Image Diffusion ModelsAuthors: Sicheng Li, Keqiang Sun, Zhixin Lai, Xiaoshi Wu, Feng Qiu, Haoran Xie, Kazunori Miyata, Hongsheng LiSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [47] arXiv:2403.18407 [pdf, other]
-
Title: A Channel-ensemble Approach: Unbiased and Low-variance Pseudo-labels is Critical for Semi-supervised ClassificationSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [48] arXiv:2403.18406 [pdf, other]
-
Title: An Image Grid Can Be Worth a Video: Zero-shot Video Question Answering Using a VLMComments: Our code is available at this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG)
- [49] arXiv:2403.18397 [pdf, ps, other]
-
Title: Colour and Brush Stroke Pattern Recognition in Abstract Art using Modified Deep Convolutional Generative Adversarial NetworksComments: 28 pages, 5 tables, 7 figuresSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [50] arXiv:2403.18383 [pdf, other]
-
Title: Generative Multi-modal Models are Good Class-Incremental LearnersComments: Accepted at CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [51] arXiv:2403.18373 [pdf, other]
-
Title: BAM: Box Abstraction Monitors for Real-time OoD Detection in Object DetectionSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [52] arXiv:2403.18370 [pdf, other]
-
Title: Ship in Sight: Diffusion Models for Ship-Image Super ResolutionComments: Accepted at 2024 International Joint Conference on Neural Networks (IJCNN)Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
- [53] arXiv:2403.18361 [pdf, other]
-
Title: ViTAR: Vision Transformer with Any ResolutionAuthors: Qihang Fan, Quanzeng You, Xiaotian Han, Yongfei Liu, Yunzhe Tao, Huaibo Huang, Ran He, Hongxia YangSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [54] arXiv:2403.18360 [pdf, other]
-
Title: Learning CNN on ViT: A Hybrid Model to Explicitly Class-specific Boundaries for Domain AdaptationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [55] arXiv:2403.18356 [pdf, other]
-
Title: MonoHair: High-Fidelity Hair Modeling from a Monocular VideoAuthors: Keyu Wu, Lingchen Yang, Zhiyi Kuang, Yao Feng, Xutao Han, Yuefan Shen, Hongbo Fu, Kun Zhou, Youyi ZhengComments: Accepted by IEEE CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [56] arXiv:2403.18351 [pdf, other]
-
Title: Generating Diverse Agricultural Data for Vision-Based Farming ApplicationsAuthors: Mikolaj Cieslak, Umabharathi Govindarajan, Alejandro Garcia, Anuradha Chandrashekar, Torsten Hädrich, Aleksander Mendoza-Drosik, Dominik L. Michels, Sören Pirk, Chia-Chun Fu, Wojciech PałubickiComments: 10 pages, 8 figures, 3 tablesSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Graphics (cs.GR); Machine Learning (cs.LG)
- [57] arXiv:2403.18342 [pdf, other]
-
Title: Learning Inclusion Matching for Animation Paint Bucket ColorizationComments: accepted to CVPR 2024. Project Page: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [58] arXiv:2403.18334 [pdf, other]
-
Title: DODA: Diffusion for Object-detection Domain Adaptation in AgricultureSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [59] arXiv:2403.18330 [pdf, other]
-
Title: Tracking-Assisted Object Detection with Event CamerasAuthors: Ting-Kang Yen, Igor Morawski, Shusil Dangi, Kai He, Chung-Yi Lin, Jia-Fong Yeh, Hung-Ting Su, Winston HsuSubjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
- [60] arXiv:2403.18328 [pdf, other]
-
Title: PIPNet3D: Interpretable Detection of Alzheimer in MRI ScansAuthors: Lisa Anita De Santi, Jörg Schlötterer, Michael Scheschenja, Joel Wessendorf, Meike Nauta, Vincenzo Positano, Christin SeifertSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [61] arXiv:2403.18318 [pdf, other]
-
Title: Uncertainty-Aware SAR ATR: Defending Against Adversarial Attacks via Bayesian Neural NetworksSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [62] arXiv:2403.18294 [pdf, other]
-
Title: Multi-scale Unified Network for Image ClassificationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [63] arXiv:2403.18293 [pdf, other]
-
Title: Efficient Test-Time Adaptation of Vision-Language ModelsComments: Accepted to CVPR 2024. The code has been released in \url{this https URL}Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [64] arXiv:2403.18291 [pdf, other]
-
Title: Towards Non-Exemplar Semi-Supervised Class-Incremental LearningSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [65] arXiv:2403.18282 [pdf, other]
-
Title: SGDM: Static-Guided Dynamic Module Make Stronger Visual ModelsComments: 16 pages, 4 figuresSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [66] arXiv:2403.18281 [pdf, other]
-
Title: AIR-HLoc: Adaptive Image Retrieval for Efficient Visual LocalisationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [67] arXiv:2403.18274 [pdf, other]
-
Title: DVLO: Deep Visual-LiDAR Odometry with Local-to-Global Feature Fusion and Bi-Directional Structure AlignmentSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [68] arXiv:2403.18271 [pdf, other]
-
Title: Unleashing the Potential of SAM for Medical Adaptation via Hierarchical DecodingComments: CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [69] arXiv:2403.18270 [pdf, other]
-
Title: Image Deraining via Self-supervised Reinforcement LearningSubjects: Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
- [70] arXiv:2403.18260 [pdf, other]
-
Title: Toward Interactive Regional Understanding in Vision-Large Language ModelsComments: NAACL 2024 Main ConferenceSubjects: Computer Vision and Pattern Recognition (cs.CV); Computation and Language (cs.CL)
- [71] arXiv:2403.18258 [pdf, other]
-
Title: Enhancing Generative Class Incremental Learning Performance with Model Forgetting ApproachSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [72] arXiv:2403.18252 [pdf, other]
-
Title: Beyond Embeddings: The Promise of Visual Table in Multi-Modal ModelsComments: Project page: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG); Multimedia (cs.MM)
- [73] arXiv:2403.18241 [pdf, other]
-
Title: NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and GenerationAuthors: Ruikai Cui, Weizhe Liu, Weixuan Sun, Senbo Wang, Taizhang Shang, Yang Li, Xibin Song, Han Yan, Zhennan Wu, Shenzhou Chen, Hongdong Li, Pan JiSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Graphics (cs.GR); Machine Learning (cs.LG)
- [74] arXiv:2403.18238 [pdf, other]
-
Title: TAFormer: A Unified Target-Aware Transformer for Video and Motion Joint Prediction in Aerial ScenesAuthors: Liangyu Xu, Wanxuan Lu, Hongfeng Yu, Yongqiang Mao, Hanbo Bi, Chenglong Liu, Xian Sun, Kun FuComments: 17 pages, 9 figuresSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [75] arXiv:2403.18228 [pdf, other]
-
Title: Fourier or Wavelet bases as counterpart self-attention in spikformer for efficient visual classificationComments: 18 pages, 2 figures. arXiv admin note: substantial text overlap with arXiv:2308.02557Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE)
- [76] arXiv:2403.18211 [pdf, other]
-
Title: NeuroPictor: Refining fMRI-to-Image Reconstruction via Multi-individual Pretraining and Multi-level ModulationSubjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
- [77] arXiv:2403.18208 [pdf, other]
-
Title: An Evolutionary Network Architecture Search Framework with Adaptive Multimodal Fusion for Hand Gesture RecognitionSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Neural and Evolutionary Computing (cs.NE)
- [78] arXiv:2403.18207 [pdf, other]
-
Title: Road Obstacle Detection based on Unknown Objectness ScoresComments: ICRA 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
- [79] arXiv:2403.18201 [pdf, other]
-
Title: Few-shot Online Anomaly Detection and SegmentationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [80] arXiv:2403.18193 [pdf, other]
-
Title: Middle Fusion and Multi-Stage, Multi-Form Prompts for Robust RGB-T TrackingSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [81] arXiv:2403.18187 [pdf, other]
-
Title: LayoutFlow: Flow Matching for Layout GenerationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [82] arXiv:2403.18186 [pdf, other]
-
Title: Don't Look into the Dark: Latent Codes for Pluralistic Image InpaintingComments: cvpr 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [83] arXiv:2403.18180 [pdf, other]
-
Title: Multi-Layer Dense Attention Decoder for Polyp SegmentationSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [84] arXiv:2403.18158 [pdf, other]
-
Title: The Effects of Short Video-Sharing Services on Video Copy DetectionSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [85] arXiv:2403.18118 [pdf, other]
-
Title: EgoLifter: Open-world 3D Segmentation for Egocentric PerceptionComments: Preprint. Project page: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [86] arXiv:2403.18117 [pdf, ps, other]
-
Title: TDIP: Tunable Deep Image Processing, a Real Time Melt Pool Monitoring SolutionSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [87] arXiv:2403.18116 [pdf, other]
-
Title: QuakeSet: A Dataset and Low-Resource Models to Monitor Earthquakes through Sentinel-1Comments: Accepted at ISCRAM 2024Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
- [88] arXiv:2403.18114 [pdf, other]
-
Title: Segment Any Medical Model ExtendedAuthors: Yihao Liu, Jiaming Zhang, Andres Diaz-Pinto, Haowei Li, Alejandro Martin-Gomez, Amir Kheradmand, Mehran ArmandComments: The content of the manuscript has been presented in SPIE Medical Imaging 2024, and had been accepted to appear in the proceedings of the conferenceSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [89] arXiv:2403.18104 [pdf, other]
-
Title: Mathematical Foundation and Corrections for Full Range Head Pose EstimationSubjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
- [90] arXiv:2403.18094 [pdf, other]
-
Title: A Personalized Video-Based Hand Taxonomy: Application for Individuals with Spinal Cord InjurySubjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
- [91] arXiv:2403.18092 [pdf, other]
-
Title: OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware InterpolationComments: CVPR 2024Subjects: Computer Vision and Pattern Recognition (cs.CV)
- [92] arXiv:2403.18080 [pdf, other]
-
Title: EgoPoseFormer: A Simple Baseline for Egocentric 3D Human Pose EstimationAuthors: Chenhongyi Yang, Anastasia Tkach, Shreyas Hampali, Linguang Zhang, Elliot J. Crowley, Cem KeskinComments: Tech ReportSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [93] arXiv:2403.18074 [pdf, other]
-
Title: Every Shot Counts: Using Exemplars for Repetition Counting in VideosComments: Project website: this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
- [94] arXiv:2403.18067 [pdf, other]
-
Title: State of the art applications of deep learning within tracking and detecting marine debris: A surveyComments: Review paper, 60 pages including references, 1 figure, 3 tables, 1 supplementary dataSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [95] arXiv:2403.18063 [pdf, other]
-
Title: Spectral Convolutional Transformer: Harmonizing Real vs. Complex Multi-View Spectral Operators for Vision TransformerSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG); Multimedia (cs.MM)
- [96] arXiv:2403.18040 [pdf, other]
-
Title: Global Point Cloud Registration Network for Large TransformationsAuthors: Hanz Cuevas-Velasquez, Alejandro Galán-Cuenca, Antonio Javier Gallego, Marcelo Saval-Calvo, Robert B. FisherSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [97] arXiv:2403.18038 [pdf, ps, other]
-
Title: TGGLinesPlus: A robust topological graph-guided computer vision algorithm for line detection from imagesComments: Our TGGLinesPlus Python implementation is open source. 27 pages, 8 figures and 4 tablesSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [98] arXiv:2403.18036 [pdf, other]
-
Title: Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene AffordanceAuthors: Zan Wang, Yixin Chen, Baoxiong Jia, Puhao Li, Jinlu Zhang, Jingze Zhang, Tengyu Liu, Yixin Zhu, Wei Liang, Siyuan HuangComments: CVPR 2024; 16 pagesSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [99] arXiv:2403.18033 [pdf, other]
-
Title: SpectralWaste Dataset: Multimodal Data for Waste Sorting AutomationAuthors: Sara Casao, Fernando Peña, Alberto Sabater, Rosa Castillón, Darío Suárez, Eduardo Montijano, Ana C. MurilloSubjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
- [100] arXiv:2403.17998 [pdf, other]
-
Title: Text Is MASS: Modeling as Stochastic Embedding for Text-Video RetrievalAuthors: Jiamian Wang, Guohao Sun, Pichao Wang, Dongfang Liu, Sohail Dianat, Majid Rabbani, Raghuveer Rao, Zhiqiang TaoComments: Accepted by CVPR 2024, code and model are available at this https URLSubjects: Computer Vision and Pattern Recognition (cs.CV)
- [101] arXiv:2403.17995 [pdf, other]
-
Title: Semi-Supervised Image Captioning Considering Wasserstein Graph MatchingAuthors: Yang YangSubjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
- [102] arXiv:2403.17994 [pdf, other]
-
Title: Solution for Point Tracking Task of ICCV 1st Perception Test Challenge 2023Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
- [103] arXiv:2403.18821 (cross-list from cs.SD) [pdf, other]
-
Title: Real Acoustic Fields: An Audio-Visual Room Acoustics Dataset and BenchmarkAuthors: Ziyang Chen, Israel D. Gebru, Christian Richardt, Anurag Kumar, William Laney, Andrew Owens, Alexander RichardComments: Accepted to CVPR 2024. Project site: this https URLSubjects: Sound (cs.SD); Computer Vision and Pattern Recognition (cs.CV); Multimedia (cs.MM); Audio and Speech Processing (eess.AS)
- [104] arXiv:2403.18734 (cross-list from eess.IV) [pdf, other]
-
Title: A vascular synthetic model for improved aneurysm segmentation and detection via Deep Neural NetworksSubjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[ showing 104 entries per page: fewer | more | all ]
Disable MathJax (What is MathJax?)
Links to: arXiv, form interface, find, cs, new, 2403, contact, help (Access key information)