2231 papers
cs cv
0/02026
49- JulGaussianFusion: Unified 3D Gaussian Representation for Multi-Modal Fusion Perceptionno summary yetcs-cv2607.00746Tencent0 citesJul 1, 2026
- JulPatch Knowledge Transfer for Efficient AI-Generated Image Quality Assessmentno summary yetcs-cv2607.05605Princeton0 citesJul 6, 2026
- JulAssociation Restoration Test: Revealing Restorable Shortcuts after Unlearningno summary yetcs-cv2607.05726Stanford0 citesJul 7, 2026
- JulCardiac MRI Through-Plane Super-Resolution Guided by Reference and Memoryno summary yetcs-cv2607.07581NYU0 citesJul 8, 2026
- JulAPIVOT: Adaptive Planning with Interleaved Vision-Language Thoughtsno summary yetcs-cv2607.08024Stanford0 citesJul 9, 2026
- JulLUMI: Tokenizer-Agnostic LLM-Based Lossless Image Compressionno summary yetcs-cv2607.08221Princeton0 citesJul 9, 2026
- JulWan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generationno summary yetcs-cv2607.09581Alibaba0 citesJul 10, 2026
- JunMAOAM: Unified Object and Material Selection with Vision-Language Modelsno summary yetcs-cv2606.04880Adobe0 citesJun 2, 2026
- JunDiffusion Transformer World-Action Model for AV Scene Predictionno summary yetcs-cv2606.12987Stanford0 citesJun 11, 2026
- JunRT-VLA: Real-Time Vision-Language-Action Models via Knowledge Distillationno summary yetcs-cv2606.14010CMU0 citesJun 12, 2026
- JunDiffusion-Refined Segmentation and Vision-Language Interpretation for Pediatric Brain Tumor MRIno summary yetcs-cv2606.14072Stanford0 citesJun 12, 2026
- JunGiving AI a Headache: Acoustic Adversarial Attacks to Computer Vision Applicationsno summary yetcs-cv2606.14658CMU0 citesJun 12, 2026
- JunSelf-Questioning Vision-Language Models: Reinforcement Learning for Compositional Visual Reasoningno summary yetcs-cv2606.15651MIT0 citesJun 14, 2026
- JunDecoupling Semantics from Distortions: Multi-Scale Two-Stream Vision-Language Alignment for AI-Generated Image Quality Assessmentno summary yetcs-cv2606.16799Princeton0 citesJun 15, 2026
- JunNeural Tree Reconstruction for the Open Forest Observatoryno summary yetcs-cv2606.18153Berkeley0 citesJun 16, 2026
- JunSpectralDiT: Timestep-Conditioned Spectral Residual Correction for Flow-Matching DiTsno summary yetcs-cv2606.18765Princeton0 citesJun 17, 2026
- JunParaScale: Scale-Calibrated Camera-Motion Transfer via a Gauge-Invariant Parallax Numberno summary yetcs-cv2606.19805Princeton0 citesJun 18, 2026
- JunShear-Free Viewport Magnification for 360-Degree via Spherical Mobius Boostsno summary yetcs-cv2606.20684Princeton0 citesJun 14, 2026
- JunTranslating Inference-Time Control to Radiology Vision-Language Models: Activation Steering for Pneumonia Classification on Chest X-raysno summary yetcs-cv2606.20852Stanford0 citesJun 18, 2026
- JunSCOPE: Scale-Consistent One-Pass Estimation of 3D Geometryno summary yetcs-cv2606.21300Alibaba0 citesJun 19, 2026
- JunZero-Shot Vision-Language Models for Classroom Engagement Recognition: A Benchmark Study of Prompt Sensitivity and Cross-Dataset Generalizationno summary yetcs-cv2606.21861CMU0 citesJun 20, 2026
- JunDual-Stream EEG Decoding for 3D Visual Perceptionno summary yetcs-cv2606.22182MIT0 citesJun 20, 2026
- JunOrthoMotion:Disentangling Camera and Subject Motion via Geometry Semantics Orthogonal Attentionno summary yetcs-cv2606.22835Princeton0 citesJun 22, 2026
- JunUnlimited OCR Worksno summary yetcs-cv2606.23050Baidu0 citesJun 22, 2026
- JunRT-DocLayout: Real-Time End-to-End Document Layout Analysis with Reading Order in the Wildno summary yetcs-cv2606.23344Baidu0 citesJun 22, 2026
- JunGeoFidelity-Bench: Evaluating Segment-Level Geographic Fidelity in Text-to-Image Street-View Generationno summary yetcs-cv2606.23669CMU0 citesJun 22, 2026
- JunLift4D: Harmonizing Single-View 3D Estimation for 4D Reconstruction In-the-Wildno summary yetcs-cv2606.23688CMU0 citesJun 22, 2026
- JunPatternGSL: A Structured Specification Language for Template-Free and Simulation-Ready 3D Garmentsno summary yetcs-cv2606.24564CMU0 citesJun 23, 2026
- JunDiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cuesno summary yetcs-cv2606.26602Princeton0 citesJun 25, 2026
- JunAnimation2Code: Evaluating Temporal Visual Reasoning in Video-to-Code Generationno summary yetcs-cv2606.28593Berkeley0 citesJun 26, 2026
- JunEarly Warning Signals for OpenVLA Failure under Visual Distribution Shiftno summary yetcs-cv2606.29699NYU0 citesJun 29, 2026
- JunSimple Supervision Is Hard to Beat: A Bitter Lesson from Sparse Target Labels in Domain-Adaptive Object Detectionno summary yetcs-cv2606.30795Amazon0 citesJun 29, 2026
- JunPlanar-SfM: Camera Pose Estimation via Homography Graph Embeddingsno summary yetcs-cv2606.31979Amazon0 citesJun 30, 2026
- JunLearning 3D Affordances for Blade Insertion in Cluttered Stowingno summary yetcs-cv2607.02549Amazon0 citesJun 25, 2026
- JunReliability-Aware Monocular Depth Supervision for Sparse-View Neural Reconstructionno summary yetcs-cv2607.02554Stanford0 citesJun 27, 2026
- MayVideo Generation with Predictive Latentsno summary yetcs-cv2605.02134ByteDance SeedMay 4, 2026
- MayD-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Modelsno summary yetcs-cv2605.05204Tongyi-MAIMay 6, 2026
- MayWhat Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusionno summary yetcs-cv2605.07915alibaba-incMay 8, 2026
- MayCollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Modelsno summary yetcs-cv2605.08735KAIST AIMay 9, 2026
- MayR-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flowno summary yetcs-cv2605.13838Tencent0 citesMay 13, 2026
- MayFashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customizationno summary yetcs-cv2605.15824alibaba-incMay 15, 2026
- MayUnlocking Dense Metric Depth Estimation in VLMsno summary yetcs-cv2605.15876Tencent HunyuanMay 15, 2026
- MayBodyReLux: Temporally Consistent Full-Body Video Relightingno summary yetcs-cv2605.21766Netflix0 citesMay 20, 2026
- MayBernini: Latent Semantic Planning for Video Diffusionno summary yetcs-cv2605.22344ByteDanceMay 21, 2026
- MayWorldKV: Efficient World Memory with World Retrieval and Compressionno summary yetcs-cv2605.22718KAIST AIMay 21, 2026
- MayEgoRelight: Egocentric Human Capture and Illumination Recovery for Relightable and Photoreal Avatar Renderingno summary yetcs-cv2605.28401Google Research0 citesMay 27, 2026
- AprMedP-CLIP: Medical CLIP with Region-Aware Prompt Integrationno summary yetcs-cv2604.11197Alibaba0 citesApr 13, 2026
- FebObject-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detectionno summary yetcs-cv2602.20627Amazon0 citesFeb 24, 2026
- JanFreeOrbit4D: Training-Free Arbitrary Camera Redirection for Monocular Videos via Foreground-Complete 4D Reconstructionno summary yetcs-cv2601.18993Netflix0 citesJan 26, 2026
2024
55- DecMVImgNet2.0: A Larger-scale Dataset of Multi-view Imagesno summary yetcs-cv2412.01430Alibaba1 citesDec 2, 2024
- DecFitting Spherical Gaussians to Dynamic HDRI Sequencesno summary yetcs-cv2412.06511Netflix0 citesDec 9, 2024
- DecSoftPatch+: Fully Unsupervised Anomaly Classification and Segmentationno summary yetcs-cv2412.20870Tencent14 citesDec 30, 2024
- OctCafca: High-quality Novel View Synthesis of Expressive Faces from Casual Few-shot Capturesno summary yetcs-cv2410.00630Google Research11 citesOct 1, 2024
- OctDifFRelight: Diffusion-Based Facial Performance Relightingno summary yetcs-cv2410.08188Netflix13 citesOct 10, 2024
- OctLeveraging Semantic Cues from Foundation Vision Models for Enhanced Local Feature Correspondenceno summary yetcs-cv2410.09533Microsoft Research0 citesOct 12, 2024
- OctHairmony: Fairness-aware hairstyle classificationno summary yetcs-cv2410.11528Microsoft Research4 citesOct 15, 2024
- OctAnalysis and Benchmarking of Extending Blind Face Image Restoration to Videosno summary yetcs-cv2410.11828Tencent4 citesOct 15, 2024
- OctAttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Modelsno summary yetcs-cv2410.16820Alibaba4 citesOct 22, 2024
- OctOffline Evaluation of Set-Based Text-to-Image Generationno summary yetcs-cv2410.17331Google Research1 citesOct 22, 2024
- SepLow-Resolution Object Recognition with Cross-Resolution Relational Contrastive Distillationno summary yetcs-cv2409.02555Baidu23 citesSep 4, 2024
- SepPIP: Detecting Adversarial Examples in Large Vision-Language Models via Attention Patterns of Irrelevant Probe Questionsno summary yetcs-cv2409.05076Tencent5 citesSep 8, 2024
- SepDecoupling Contact for Fine-Grained Motion Style Transferno summary yetcs-cv2409.05387Tencent5 citesSep 9, 2024
- SepA Diffusion Approach to Radiance Field Relighting using Multi-Illumination Synthesisno summary yetcs-cv2409.0894714 citesSep 13, 2024score 1
- SepLVCD: Reference-based Lineart Video Colorization with Diffusion Modelsno summary yetcs-cv2409.1296011 citesSep 19, 2024score 2
- SepEnd to End Face Reconstruction via Differentiable PnPno summary yetcs-cv2409.14249Tencent0 citesSep 21, 2024
- SepNeural Light Spheres for Implicit Image Stitching and View Synthesisno summary yetcs-cv2409.17924Google Research3 citesSep 26, 2024
- SepNeural Product Importance Sampling via Warp Compositionno summary yetcs-cv2409.18974Adobe3 citesSep 12, 2024
- SepSpaceMesh: A Continuous Representation for Learning Manifold Surface Meshesno summary yetcs-cv2409.20562NVIDIA7 citesSep 30, 2024
- SepRoMo: A Robust Solver for Full-body Unlabeled Optical Motion Captureno summary yetcs-cv2410.02788Tencent1 citesSep 18, 2024
- AugMDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generationno summary yetcs-cv2408.03312Tencent5 citesAug 6, 2024
- AugFlexible 3D Lane Detection by Hierarchical Shape MatchingFlexible 3D Lane Detection by Hierarchical Shape Matchingno summary yetcs-cv2408.07163Tencent3 citesAug 13, 2024
- AugToward a More Complete OMR Solutionno summary yetcs-cv2409.00316AllenAI0 citesAug 31, 2024
- JulfVDB: A Deep-Learning Framework for Sparse, Large-Scale, and High-Performance Spatial Intelligenceno summary yetcs-cv2407.01781NVIDIA18 citesJul 1, 2024
- JulNeural varifolds: an aggregate representation for quantifying the geometry of point cloudsno summary yetcs-cv2407.04844Meta / FAIR0 citesJul 5, 2024
- JulUnderstanding Visual Feature Reliance through the Lens of Complexityno summary yetcs-cv2407.060760 citesJul 8, 2024score 2
- JulPattern Guided UV Recovery for Realistic Video Garment Texturingno summary yetcs-cv2407.10137Adobe0 citesJul 14, 2024
- JulLite2Relight: 3D-aware Single Image Portrait Relightingno summary yetcs-cv2407.10487Google Research11 citesJul 15, 2024
- JulEmoFace: Audio-driven Emotional 3D Face Animationno summary yetcs-cv2407.12501Tencent12 citesJul 17, 2024
- JulAdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detectionno summary yetcs-cv2407.15795Tencent78 citesJul 22, 2024
- JulDAC: 2D-3D Retrieval with Noisy Labels via Divide-and-Conquer Alignment and Correctionno summary yetcs-cv2407.17779Tencent1 citesJul 25, 2024
- JulLearning Spectral-Decomposed Tokens for Domain Generalized Semantic Segmentationno summary yetcs-cv2407.18568Tencent15 citesJul 26, 2024
- JulMatting by Generationno summary yetcs-cv2407.210173 citesJul 30, 2024score 1
- JunDuMapNet: An End-to-End Vectorization System for City-Scale Lane-Level Map Generationno summary yetcs-cv2406.14255Baidu6 citesJun 20, 2024
- JunDynamic Gaussian Marbles for Novel View Synthesis of Casual Monocular Videosno summary yetcs-cv2406.18717Google Research24 citesJun 26, 2024
- JunSimplicits: Mesh-Free, Geometry-Agnostic, Elastic Simulationno summary yetcs-cv2407.09497NVIDIA14 citesJun 9, 2024
- MayRGB$\leftrightarrow$X: Image decomposition and synthesis using material- and lighting-aware diffusion modelsno summary yetcs-cv2405.00666Adobe40 citesMay 1, 2024
- MayRip-NeRF: Anti-aliasing Radiance Fields with Ripmap-Encoded Platonic Solidsno summary yetcs-cv2405.02386Tencent14 citesMay 3, 2024
- MayPhysics-based Scene Layout Generation from Human Motionno summary yetcs-cv2405.12460Tencent3 citesMay 21, 2024
- MayWearable-based behaviour interpolation for semi-supervised human activity recognitionno summary yetcs-cv2405.15962Tencent15 citesMay 24, 2024
- MayGaussianPrediction: Dynamic 3D Gaussian Prediction for Motion Extrapolation and Free View Synthesisno summary yetcs-cv2405.19745Google Research9 citesMay 30, 2024
- MayFiltering After Shading With Stochastic Texture Filteringno summary yetcs-cv2407.06107NVIDIA6 citesMay 14, 2024
- AprFaceFolds: Meshed Radiance Manifolds for Efficient Volumetric Rendering of Dynamic Facesno summary yetcs-cv2404.13807Google Research2 citesApr 22, 2024
- AprGaussianTalker: Speaker-specific Talking Head Synthesis via 3D Gaussian Splattingno summary yetcs-cv2404.14037Alibaba22 citesApr 22, 2024
- AprXFeat: Accelerated Features for Lightweight Image Matchingno summary yetcs-cv2404.19174Microsoft Research3 citesApr 30, 2024
- MarDual-path Frequency Discriminators for Few-shot Anomaly Detectionno summary yetcs-cv2403.04151Tencent12 citesMar 7, 2024
- MarUnsupervised Modality-Transferable Video Highlight Detection with Representation Activation Sequence Learningno summary yetcs-cv2403.09401Tencent7 citesMar 14, 2024
- MarLightIt: Illumination Modeling and Control for Diffusion Modelsno summary yetcs-cv2403.106152 citesMar 15, 2024score 5
- MarSelf-Supervised Learning for Medical Image Data with Anatomy-Oriented Imaging Planesno summary yetcs-cv2403.16499Tencent11 citesMar 25, 2024
- MarFastPerson: Enhancing Video Learning through Effective Video Summarization that Preserves Linguistic and Visual Contextsno summary yetcs-cv2403.17727Sony10 citesMar 26, 2024
- MarMetric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimationno summary yetcs-cv2404.15506Tencent152 citesMar 22, 2024
- FebVirtual Classification: Modulating Domain-Specific Knowledge for Multidomain Crowd Countingno summary yetcs-cv2402.03758Alibaba12 citesFeb 6, 2024
- FebEnhancing Hyperspectral Images via Diffusion Model and Group-Autoencoder Super-resolution Networkno summary yetcs-cv2402.17285Alibaba14 citesFeb 27, 2024
- FebEnhancing Visual Document Understanding with Contrastive Learning in Large Visual-Language Modelsno summary yetcs-cv2402.19014Tencent26 citesFeb 29, 2024
- JanMulti-Track Timeline Control for Text-Driven 3D Human Motion Generationno summary yetcs-cv2401.08559NVIDIA0 citesJan 16, 2024
2023
57- NovLearning Discriminative Features for Crowd Countingno summary yetcs-cv2311.04509Baidu1 citesNov 8, 2023
- NovSynthetically Enhanced: Unveiling Synthetic Data's Potential in Medical Imaging Researchno summary yetcs-cv2311.09402Google Research2 citesNov 15, 2023
- OctMUSCLE: Multi-task Self-supervised Continual Learning to Pre-train Deep Models for X-ray Images of Multiple Body Partsno summary yetcs-cv2310.02000Baidu19 citesOct 3, 2023
- OctDUSA: Decoupled Unsupervised Sim2Real Adaptation for Vehicle-to-Everything Collaborative Perceptionno summary yetcs-cv2310.08117Baidu16 citesOct 12, 2023
- OctIs ImageNet worth 1 video? Learning strong image encoders from 1 long unlabelled videono summary yetcs-cv2310.08584DeepMind0 citesOct 12, 2023
- OctWoodpecker: Hallucination Correction for Multimodal Large Language Modelsno summary yetcs-cv2310.1604522 citesOct 24, 2023score 6
- SepStroke-based Neural Painting and Stylization with Dynamically Predicted Painting Regionno summary yetcs-cv2309.03504Tencent18 citesSep 7, 2023
- SepToward High Quality Facial Representation Learningno summary yetcs-cv2309.03575Tencent2 citesSep 7, 2023
- SepTowards Better Multi-modal Keyphrase Generation via Visual Entity Enhancement and Multi-granularity Image Noise Filteringno summary yetcs-cv2309.04734Tencent2 citesSep 9, 2023
- SepSoccerNet 2023 Challenges Resultsno summary yetcs-cv2309.06006Baidu1 citesSep 12, 2023
- SepNucleus-aware Self-supervised Pretraining Using Unpaired Image-to-image Translation for Histopathology Imagesno summary yetcs-cv2309.07394Alibaba12 citesSep 14, 2023
- SepAdSEE: Investigating the Impact of Image Style Editing on Advertisement Attractivenessno summary yetcs-cv2309.08159Tencent0 citesSep 15, 2023
- SepTowards Real-Time Neural Video Codec for Cross-Platform Application Using Calibration Informationno summary yetcs-cv2309.11276Tencent6 citesSep 20, 2023
- SepPreface: A Data-driven Volumetric Prior for Few-shot Ultra High-resolution Face Synthesisno summary yetcs-cv2309.16859Google Research1 citesSep 28, 2023
- SepAdvancing The Rate-Distortion-Computation Frontier For Neural Image Compressionno summary yetcs-cv2311.12821Google Research8 citesSep 26, 2023
- AugPVG: Progressive Vision Graph for Vision Recognitionno summary yetcs-cv2308.00574Tencent15 citesAug 1, 2023
- AugComputational Long Exposure Mobile Photographyno summary yetcs-cv2308.013790 citesAug 2, 2023score 2
- AugDiffusion-Augmented Depth Prediction with Sparse Annotationsno summary yetcs-cv2308.02283Adobe7 citesAug 4, 2023
- AugLanguage-based Photo Color Adjustment for Graphic Designsno summary yetcs-cv2308.03059Adobe49 citesAug 6, 2023
- AugMirror-NeRF: Learning Neural Radiance Fields for Mirrors with Whitted-Style Ray Tracingno summary yetcs-cv2308.0328030 citesAug 7, 2023score 2
- AugTextPainter: Multimodal Text Image Generation with Visual-harmony and Text-comprehension for Poster Designno summary yetcs-cv2308.04733Alibaba6 citesAug 9, 2023
- AugFastLLVE: Real-Time Low-Light Video Enhancement with Intensity-Aware Lookup Tableno summary yetcs-cv2308.06749Alibaba13 citesAug 13, 2023
- AugDS-Depth: Dynamic and Static Depth Estimation via a Fusion Cost Volumeno summary yetcs-cv2308.07225Tencent27 citesAug 14, 2023
- AugSelf-Supervised Learning for Endoscopic Video Analysisno summary yetcs-cv2308.12394Google Research0 citesAug 23, 2023
- AugMultiCapCLIP: Auto-Encoding Prompts for Zero-Shot Multilingual Visual Captioningno summary yetcs-cv2308.13218Tencent11 citesAug 25, 2023
- AugMVMR: A New Framework for Evaluating Faithfulness of Video Moment Retrieval against Multiple Distractorsno summary yetcs-cv2309.16701Adobe0 citesAug 15, 2023
- JulLEDITS: Real Image Editing with DDPM Inversion and Semantic Guidanceno summary yetcs-cv2307.005228 citesJul 2, 2023score 3
- JunTaming Reversible Halftoning via Predictive Luminanceno summary yetcs-cv2306.08309Tencent1 citesJun 14, 2023
- JunM3PT: A Multi-Modal Model for POI Taggingno summary yetcs-cv2306.10079Alibaba3 citesJun 16, 2023
- JunRSMT: Real-time Stylized Motion Transition for Charactersno summary yetcs-cv2306.11970Tencent17 citesJun 21, 2023
- JunHierarchical Matching and Reasoning for Multi-Query Image Retrievalno summary yetcs-cv2306.14460Baidu1 citesJun 26, 2023
- JunTraining-Free Neural Matte Extraction for Visual Effectsno summary yetcs-cv2306.17321Google Research1 citesJun 29, 2023
- MayPRSeg: A Lightweight Patch Rotate MLP Decoder for Semantic Segmentationno summary yetcs-cv2305.00671Baidu13 citesMay 1, 2023
- MayCALM: Conditional Adversarial Latent Models for Directable Virtual Charactersno summary yetcs-cv2305.02195NVIDIA51 citesMay 2, 2023
- MaySingle-Shot Implicit Morphable Faces with Consistent Texture Parameterizationno summary yetcs-cv2305.0304314 citesMay 4, 2023score 1
- MayFashionTex: Controllable Virtual Try-on with Text and Textureno summary yetcs-cv2305.04451Adobe16 citesMay 8, 2023
- MayCLIP-VG: Self-paced Curriculum Adapting of CLIP for Visual Groundingno summary yetcs-cv2305.08685Alibaba1 citesMay 15, 2023
- MayDrag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifoldno summary yetcs-cv2305.1097311 citesMay 18, 2023score 2
- MayPhotoMat: A Material Generator Learned from Single Flash Photosno summary yetcs-cv2305.12296Adobe30 citesMay 20, 2023
- AprEnhancing Deformable Local Features by Jointly Learning to Detect and Describe Keypointsno summary yetcs-cv2304.00583Google Research0 citesApr 2, 2023
- AprInstance-Aware Domain Generalization for Face Anti-Spoofingno summary yetcs-cv2304.05640Tencent5 citesApr 12, 2023
- AprLA3: Efficient Label-Aware AutoAugmentno summary yetcs-cv2304.10310Tencent1 citesApr 20, 2023
- AprAutoNeRF: Training Implicit Scene Representations with Autonomous Agentsno summary yetcs-cv2304.11241Mistral5 citesApr 21, 2023
- AprUniNeXt: Exploring A Unified Architecture for Vision Recognitionno summary yetcs-cv2304.13700Alibaba13 citesApr 26, 2023
- MarTo Make Yourself Invisible with Adversarial Semantic Contoursno summary yetcs-cv2303.00284Alibaba2 citesMar 1, 2023
- MarIterative Few-shot Semantic Segmentation from Image Label Textno summary yetcs-cv2303.05646Tencent18 citesMar 10, 2023
- MarDiffusion-based Document Layout Generationno summary yetcs-cv2303.10787Microsoft Research0 citesMar 19, 2023
- FebUnderstanding metric-related pitfalls in image analysis validationno summary yetcs-cv2302.01790NVIDIA20 citesFeb 3, 2023
- FebPolynomial Neural Fields for Subband Decomposition and Manipulationno summary yetcs-cv2302.04862Google Research8 citesFeb 9, 2023
- FebMake Your Brief Stroke Real and Stereoscopic: 3D-Aware Simplified Sketch to Portrait Generationno summary yetcs-cv2302.06857Baidu2 citesFeb 14, 2023
- FebSTOA-VLP: Spatial-Temporal Modeling of Object and Action for Video-Language Pre-trainingno summary yetcs-cv2302.09736Tencent1 citesFeb 20, 2023
- FebSelf-supervised learning of Split Invariant Equivariant representationsno summary yetcs-cv2302.10283Meta / FAIR2 citesFeb 14, 2023
- FebVid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioningno summary yetcs-cv2302.14115DeepMind15 citesFeb 27, 2023
- Jan${S}^{2}$Net: Accurate Panorama Depth Estimation on Spherical Surfaceno summary yetcs-cv2301.05845Alibaba12 citesJan 14, 2023
- JanRethinking Precision of Pseudo Label: Test-Time Adaptation via Complementary Learningno summary yetcs-cv2301.06013Tencent1 citesJan 15, 2023
- JanMulti-Camera Lighting Estimation for Photorealistic Front-Facing Mobile Augmented Realityno summary yetcs-cv2301.06143Google Research9 citesJan 15, 2023
- JanA Comprehensive Review of Modern Object Segmentation Approachesno summary yetcs-cv2301.07499Amazon41 citesJan 13, 2023
2022
78- DecVision Transformer with Attentive Pooling for Robust Facial Expression Recognitionno summary yetcs-cv2212.05463Baidu137 citesDec 11, 2022
- DecFeature Calibration Network for Occluded Pedestrian Detectionno summary yetcs-cv2212.05717Baidu34 citesDec 12, 2022
- DecHDNet: A Hierarchically Decoupled Network for Crowd Countingno summary yetcs-cv2212.05722Tencent3 citesDec 12, 2022
- DecTemporal Output Discrepancy for Loss Estimation-based Active Learningno summary yetcs-cv2212.10613Baidu10 citesDec 20, 2022
- NovCLOP: Video-and-Language Pre-Training with Knowledge Regularizationsno summary yetcs-cv2211.03314Baidu1 citesNov 7, 2022
- NovLesion Guided Explainable Few Weak-shot Medical Report Generationno summary yetcs-cv2211.08732Tencent11 citesNov 16, 2022
- NovVideogenic: Identifying Highlight Moments in Videos with Professional Photographs as a Priorno summary yetcs-cv2211.12493Adobe2 citesNov 22, 2022
- NovGlobal Meets Local: Effective Multi-Label Image Classification via Category-Aware Weak Supervisionno summary yetcs-cv2211.12716Tencent6 citesNov 23, 2022
- OctSoccerNet 2022 Challenges Resultsno summary yetcs-cv2210.02365Baidu36 citesOct 5, 2022
- OctDetaching and Boosting: Dual Engine for Scale-Invariant Self-Supervised Monocular Depth Estimationno summary yetcs-cv2210.03952Baidu2 citesOct 8, 2022
- OctTowards Understanding and Boosting Adversarial Transferability from a Distribution Perspectiveno summary yetcs-cv2210.04213Alibaba64 citesOct 9, 2022
- OctVideo Summarization Overviewno summary yetcs-cv2210.11707Microsoft Research10 citesOct 21, 2022
- OctPseudoAugment: Learning to Use Unlabeled Data for Data Augmentation in Point Cloudsno summary yetcs-cv2210.13428Google Research16 citesOct 24, 2022
- OctIncorporating Crowdsourced Annotator Distributions into Ensemble Modeling to Improve Classification Trustworthiness for Ancient Greek Papyrino summary yetcs-cv2210.16380Amazon0 citesOct 28, 2022
- SepTokenCut: Segmenting Objects in Images and Videos with Self-supervised Transformer and Normalized Cutno summary yetcs-cv2209.00383Tencent77 citesSep 1, 2022
- SepLearning to Relight Portrait Images via a Virtual Light Stage and Synthetic-to-Real Adaptationno summary yetcs-cv2209.10510NVIDIA56 citesSep 21, 2022
- SepRepsNet: Combining Vision with Language for Automated Medical Reportsno summary yetcs-cv2209.13171Google Research26 citesSep 27, 2022
- SepAdma-GAN: Attribute-Driven Memory Augmented GANs for Text-to-Image Generationno summary yetcs-cv2209.14046Tencent17 citesSep 28, 2022
- AugJoint Learning Content and Degradation Aware Feature for Blind Super-Resolutionno summary yetcs-cv2208.13436Tencent14 citesAug 29, 2022
- JulNeural Parameterization for Dynamic Human Head Editingno summary yetcs-cv2207.00210Tencent11 citesJul 1, 2022
- JulBack to MLP: A Simple Baseline for Human Motion Predictionno summary yetcs-cv2207.01567Tencent0 citesJul 4, 2022
- JulDistilling Ensemble of Explanations for Weakly-Supervised Pre-Training of Image Segmentation Modelsno summary yetcs-cv2207.03335Baidu4 citesJul 4, 2022
- JulPaint and Distill: Boosting 3D Object Detection with Semantic Passing Networkno summary yetcs-cv2207.05497Baidu11 citesJul 12, 2022
- JulBackdoor Attacks on Crowd Countingno summary yetcs-cv2207.05641Microsoft Research11 citesJul 12, 2022
- JulFactorized and Controllable Neural Re-Rendering of Outdoor Scene for Photo Extrapolationno summary yetcs-cv2207.06899Baidu18 citesJul 14, 2022
- JulDuetFace: Collaborative Privacy-Preserving Face Recognition via Channel Splitting in the Frequency Domainno summary yetcs-cv2207.07340Tencent33 citesJul 15, 2022
- JulShow Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head Attentionno summary yetcs-cv2207.08352Adobe5 citesJul 18, 2022
- JulAlignSDF: Pose-Aligned Signed Distance Fields for Hand-Object Reconstructionno summary yetcs-cv2207.12909DeepMind8 citesJul 26, 2022
- JulTowards Domain-agnostic Depth Completionno summary yetcs-cv2207.14466Adobe4 citesJul 29, 2022
- JulLess is More: Consistent Video Depth Estimation with Masked Frames Modelingno summary yetcs-cv2208.00380Adobe19 citesJul 31, 2022
- JunZero-Shot Video Question Answering via Frozen Bidirectional Language Modelsno summary yetcs-cv2206.08155DeepMind64 citesJun 16, 2022
- JunEyeNeRF: A Hybrid Representation for Photorealistic Synthesis, Animation and Relighting of Human Eyesno summary yetcs-cv2206.08428Google Research30 citesJun 16, 2022
- JunValidation of Vector Data using Oblique Imagesno summary yetcs-cv2206.09038Microsoft Research22 citesJun 17, 2022
- JunDistribution Regularized Self-Supervised Learning for Domain Adaptation of Semantic Segmentationno summary yetcs-cv2206.09683Meta / FAIR7 citesJun 20, 2022
- JunUnderstanding the effect of sparsity on neural networks robustnessno summary yetcs-cv2206.10915Google Research3 citesJun 22, 2022
- JunA Fast Text-Driven Approach for Generating Artistic Contentno summary yetcs-cv2208.01748Adobe0 citesJun 22, 2022
- MayNeural Rendering in a Room: Amodal 3D Understanding and Free-Viewpoint Rendering for the Closed Scene Composed of Pre-Captured Objectsno summary yetcs-cv2205.02714Google Research17 citesMay 5, 2022
- MayReLU Fields: The Little Non-linearity That Couldno summary yetcs-cv2205.10824Adobe75 citesMay 22, 2022
- MayJointly Optimizing Color Rendition and In-Camera Backgrounds in an RGB Virtual Production Stageno summary yetcs-cv2205.12403Netflix4 citesMay 24, 2022
- MayMultiview Textured Mesh Recovery by Differentiable Renderingno summary yetcs-cv2205.12468Alibaba21 citesMay 25, 2022
- MayNeural 3D Reconstruction in the Wildno summary yetcs-cv2205.12955Google Research102 citesMay 25, 2022
- MayPSTNet: Point Spatio-Temporal Convolution on Point Cloud Sequencesno summary yetcs-cv2205.13713Baidu69 citesMay 27, 2022
- AprDynamic Supervisor for Cross-dataset Object Detectionno summary yetcs-cv2204.00183Alibaba1 citesApr 1, 2022
- AprAligning Silhouette Topology for Self-Adaptive 3D Human Pose Recoveryno summary yetcs-cv2204.01276Google Research6 citesApr 4, 2022
- AprMGRR-Net: Multi-level Graph Relational Reasoning Network for Facial Action Units Detectionno summary yetcs-cv2204.01349Tencent3 citesApr 4, 2022
- AprNon-Local Latent Relation Distillation for Self-Adaptive 3D Human Pose Estimationno summary yetcs-cv2204.01971Google Research3 citesApr 5, 2022
- AprOn Distinctive Image Captioning via Comparing and Reweightingno summary yetcs-cv2204.03938Baidu29 citesApr 8, 2022
- AprSpatial Likelihood Voting with Self-Knowledge Distillation for Weakly Supervised Object Detectionno summary yetcs-cv2204.06899Alibaba4 citesApr 14, 2022
- AprTowards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasksno summary yetcs-cv2204.07780Tencent62 citesApr 16, 2022
- AprLearning Shape Priors by Pairwise Comparison for Robust Semantic Segmentationno summary yetcs-cv2204.11090Tencent4 citesApr 23, 2022
- AprCLIP-Art: Contrastive Pre-training for Fine-Grained Art Classificationno summary yetcs-cv2204.14244Adobe100 citesApr 29, 2022
- MarTowards Universal Backward-Compatible Representation Learningno summary yetcs-cv2203.01583Tencent0 citesMar 3, 2022
- MarTime-to-Label: Temporal Consistency for Self-Supervised Monocular 3D Object Detectionno summary yetcs-cv2203.02193Google Research0 citesMar 4, 2022
- MarDuMLP-Pin: A Dual-MLP-dot-product Permutation-invariant Network for Set Feature Extractionno summary yetcs-cv2203.04007Alibaba7 citesMar 8, 2022
- MarCAR: Class-aware Regularizations for Semantic Segmentationno summary yetcs-cv2203.07160Tencent5 citesMar 14, 2022
- MarIntegrating Language Guidance into Vision-based Deep Metric Learningno summary yetcs-cv2203.08543DeepMind1 citesMar 16, 2022
- MarObject discovery and representation networksno summary yetcs-cv2203.08777DeepMind6 citesMar 16, 2022
- MarDU-VLG: Unifying Vision-and-Language Generation via Dual Sequence-to-Sequence Pre-trainingno summary yetcs-cv2203.09052Baidu0 citesMar 17, 2022
- MarCoGS: Controllable Generation and Search from Sketch and Styleno summary yetcs-cv2203.09554Adobe1 citesMar 17, 2022
- MarUnified Line and Paragraph Detection by Graph Convolutional Networksno summary yetcs-cv2203.09638Google Research0 citesMar 17, 2022
- MarREALY: Rethinking the Evaluation of 3D Face Reconstructionno summary yetcs-cv2203.09729Tencent1 citesMar 18, 2022
- MarDomain Adaptation Meets Zero-Shot Learning: An Annotation-Efficient Approach to Multi-Modality Medical Image Segmentationno summary yetcs-cv2203.10332Tencent41 citesMar 19, 2022
- MarA Representation Separation Perspective to Correspondences-free Unsupervised 3D Point Cloud Registrationno summary yetcs-cv2203.13239Baidu8 citesMar 24, 2022
- MarIn-N-Out Generative Learning for Dense Unsupervised Video Segmentationno summary yetcs-cv2203.15312Alibaba8 citesMar 29, 2022
- MarTubeDETR: Spatio-Temporal Video Grounding with Transformersno summary yetcs-cv2203.16434DeepMind87 citesMar 30, 2022
- FebNeural Dual Contouringno summary yetcs-cv2202.01999Google Research82 citesFeb 4, 2022
- FebWebly Supervised Concept Expansion for General Purpose Vision Modelsno summary yetcs-cv2202.02317AllenAI2 citesFeb 4, 2022
- FebAmplitude Spectrum Transformation for Open Compound Domain Adaptive Semantic Segmentationno summary yetcs-cv2202.04287Google Research2 citesFeb 9, 2022
- FebOctAttention: Octree-Based Large-Scale Contexts Model for Point Cloud Compressionno summary yetcs-cv2202.06028Tencent173 citesFeb 12, 2022
- FebHierarchical Point Cloud Encoding and Decoding with Lightweight Self-Attention based Modelno summary yetcs-cv2202.06407Alibaba6 citesFeb 13, 2022
- FebSelf-Supervised Transformers for Unsupervised Object Discovery using Normalized Cutno summary yetcs-cv2202.11539Tencent143 citesFeb 23, 2022
- FebMeta-RangeSeg: LiDAR Sequence Semantic Segmentation Using Multiple Feature Aggregationno summary yetcs-cv2202.13377Alibaba50 citesFeb 27, 2022
- JanMERLOT Reserve: Neural Script Knowledge through Vision and Language and Soundno summary yetcs-cv2201.02639AllenAI9 citesJan 7, 2022
- JanMobilePhys: Personalized Mobile Camera-Based Contactless Physiological Sensingno summary yetcs-cv2201.04039Microsoft Research17 citesJan 11, 2022
- JanAction Keypoint Network for Efficient Video Recognitionno summary yetcs-cv2201.06304Baidu10 citesJan 17, 2022
- JanDeformable One-Dimensional Object Detection for Routing and Manipulationno summary yetcs-cv2201.06775Google Research28 citesJan 18, 2022
- JanReading-strategy Inspired Visual Representation Learning for Text-to-Video Retrievalno summary yetcs-cv2201.09168Alibaba75 citesJan 23, 2022
- JanModeling the Background for Incremental and Weakly-Supervised Semantic Segmentationno summary yetcs-cv2201.13338Meta / FAIR13 citesJan 31, 2022
2021
400- DecGraph Convolutional Module for Temporal Action Localization in Videosno summary yetcs-cv2112.00302Tencent0 citesDec 1, 2021
- DecControllable Video Captioning with an Exemplar Sentenceno summary yetcs-cv2112.01073Tencent20 citesDec 2, 2021
- DecDimensions of Motion: Monocular Prediction through Flow Subspacesno summary yetcs-cv2112.01502Google Research0 citesDec 2, 2021
- DecNext Day Wildfire Spread: A Machine Learning Data Set to Predict Wildfire Spreading from Remote-Sensing Datano summary yetcs-cv2112.02447Google Research122 citesDec 4, 2021
- DecDeblurring via Stochastic Refinementno summary yetcs-cv2112.02475Google Research0 citesDec 5, 2021
- Dec3D-VField: Adversarial Augmentation of Point Clouds for Domain Generalization in 3D Object Detectionno summary yetcs-cv2112.04764Google Research62 citesDec 9, 2021
- DecInterpolated Joint Space Adversarial Training for Robust and Generalizable Defensesno summary yetcs-cv2112.06323Adobe1 citesDec 12, 2021
- Dec3D Question Answeringno summary yetcs-cv2112.08359Microsoft Research0 citesDec 15, 2021
- NovEgocentric Human Trajectory Forecasting with a Wearable Camera and Multi-Modal Fusionno summary yetcs-cv2111.00993Tencent19 citesNov 1, 2021
- NovMasking Modalities for Cross-modal Video Retrievalno summary yetcs-cv2111.01300Google Research0 citesNov 1, 2021
- NovExploring the Semi-supervised Video Object Segmentation Problem from a Cyclic Perspectiveno summary yetcs-cv2111.01323Adobe0 citesNov 2, 2021
- NovBootstrap Your Object Detector via Mixed Trainingno summary yetcs-cv2111.03056Microsoft Research0 citesNov 4, 2021
- NovRecognizing Vector Graphics without Rasterizationno summary yetcs-cv2111.03281Tencent0 citesNov 5, 2021
- NovImproving Visual Quality of Image Synthesis by A Token-based Generator with Transformersno summary yetcs-cv2111.03481Microsoft Research5 citesNov 5, 2021
- NovNatural Adversarial Objectsno summary yetcs-cv2111.04204AllenAI1 citesNov 7, 2021
- NovData Augmentation Can Improve Robustnessno summary yetcs-cv2111.05328Google Research13 citesNov 9, 2021
- NovDance In the Wild: Monocular Human Animation with Neural Dynamic Appearance Synthesisno summary yetcs-cv2111.05916Adobe1 citesNov 10, 2021
- NovD$^2$LV: A Data-Driven and Local-Verification Approach for Image Copy Detectionno summary yetcs-cv2111.07090Baidu2 citesNov 13, 2021
- NovOccluded Video Instance Segmentation: Dataset and ICCV 2021 Challengeno summary yetcs-cv2111.07950Alibaba0 citesNov 15, 2021
- NovBag of Tricks and A Strong baseline for Image Copy Detectionno summary yetcs-cv2111.08004Baidu4 citesNov 13, 2021
- NovOpen Vocabulary Object Detection with Pseudo Bounding-Box Labelsno summary yetcs-cv2111.09452Salesforce7 citesNov 18, 2021
- NovBoosting Supervised Learning Performance with Co-trainingno summary yetcs-cv2111.09797NVIDIA1 citesNov 18, 2021
- NovCorrecting Face Distortion in Wide-Angle Videosno summary yetcs-cv2111.09950Google Research0 citesNov 18, 2021
- NovLOLNeRF: Learn from One Lookno summary yetcs-cv2111.09996Google Research0 citesNov 19, 2021
- NovAPANet: Adaptive Prototypes Alignment Network for Few-Shot Semantic Segmentationno summary yetcs-cv2111.12263Tencent48 citesNov 24, 2021
- NovPolyViT: Co-training Vision Transformers on Images, Videos and Audiono summary yetcs-cv2111.12993Google Research2 citesNov 25, 2021
- NovAttribute-specific Control Units in StyleGAN for Fine-grained Image Manipulationno summary yetcs-cv2111.13010Tencent13 citesNov 25, 2021
- NovVaxNeRF: Revisiting the Classic for Voxel-Accelerated Neural Radiance Fieldno summary yetcs-cv2111.13112Google Research0 citesNov 25, 2021
- NovPOEM: 1-bit Point-wise Operations based on Expectation-Maximization for Efficient Point Cloud Processingno summary yetcs-cv2111.13386Baidu3 citesNov 26, 2021
- NovUrban Radiance Fieldsno summary yetcs-cv2111.14643Google Research0 citesNov 29, 2021
- NovSearching the Search Space of Vision Transformerno summary yetcs-cv2111.14725Microsoft Research1 citesNov 29, 2021
- NovSketchEdit: Mask-Free Local Image Manipulation with Partial Sketchesno summary yetcs-cv2111.15078Adobe2 citesNov 30, 2021
- OctHighlightMe: Detecting Highlights from Human-Centric Videosno summary yetcs-cv2110.01774Adobe1 citesOct 5, 2021
- OctWaypoint Models for Instruction-guided Navigation in Continuous Environmentsno summary yetcs-cv2110.02207Microsoft Research1 citesOct 5, 2021
- OctConstruction Site Safety Monitoring and Excavator Activity Analysis Systemno summary yetcs-cv2110.03083Baidu0 citesOct 6, 2021
- OctNeural Strokes: Stylized Line Drawing of 3D Shapesno summary yetcs-cv2110.03900Adobe0 citesOct 8, 2021
- OctRobustness Evaluation of Transformer-based Form Field Extractors via Form Attacksno summary yetcs-cv2110.04413Salesforce1 citesOct 8, 2021
- OctUnsupervised Representation Learning Meets Pseudo-Label Supervised Self-Distillation: A New Approach to Rare Disease Classificationno summary yetcs-cv2110.04558Tencent12 citesOct 9, 2021
- OctVector-quantized Image Modeling with Improved VQGANno summary yetcs-cv2110.04627Google Research92 citesOct 9, 2021
- OctBuildingNet: Learning to Label 3D Buildingsno summary yetcs-cv2110.04955Adobe0 citesOct 11, 2021
- OctBoosting Fast Adversarial Training with Learnable Adversarial Initializationno summary yetcs-cv2110.05007Tencent61 citesOct 11, 2021
- OctLearning Realistic Human Reposing using Cyclic Self-Supervision with 3D Shape, Pose, and Appearance Consistencyno summary yetcs-cv2110.05458Amazon0 citesOct 11, 2021
- OctSemi-Supervised Semantic Segmentation via Adaptive Equalization Learningno summary yetcs-cv2110.05474Microsoft Research14 citesOct 11, 2021
- OctConsidering user agreement in learning to predict the aesthetic qualityno summary yetcs-cv2110.06956Tencent2 citesOct 13, 2021
- OctTowards Language-guided Visual Recognition via Dynamic Convolutionsno summary yetcs-cv2110.08797Tencent25 citesOct 17, 2021
- OctA Unified Framework for Generalized Low-Shot Medical Image Segmentation with Scarce Datano summary yetcs-cv2110.09260Tencent51 citesOct 18, 2021
- OctSTALP: Style Transfer with Auxiliary Limited Pairingno summary yetcs-cv2110.10501Adobe0 citesOct 20, 2021
- OctLearning 3D Semantic Segmentation with only 2D Image Supervisionno summary yetcs-cv2110.11325Google Research4 citesOct 21, 2021
- OctSpectrum-to-Kernel Translation for Accurate Blind Image Super-Resolutionno summary yetcs-cv2110.12151Tencent7 citesOct 23, 2021
- OctReachability Embeddings: Scalable Self-Supervised Representation Learning from Mobility Trajectories for Multimodal Geospatial Computer Visionno summary yetcs-cv2110.12521Apple6 citesOct 24, 2021
- OctImage Quality Assessment using Contrastive Learningno summary yetcs-cv2110.13266Google Research258 citesOct 25, 2021
- OctH-NeRF: Neural Radiance Fields for Rendering and Temporal Reconstruction of Humans in Motionno summary yetcs-cv2110.13746Google Research1 citesOct 26, 2021
- OctSOAT: A Scene- and Object-Aware Transformer for Vision-and-Language Navigationno summary yetcs-cv2110.14143Microsoft Research4 citesOct 27, 2021
- OctNeural-PIL: Neural Pre-Integrated Lighting for Reflectance Decompositionno summary yetcs-cv2110.14373Google Research19 citesOct 27, 2021
- OctOn-device Real-time Hand Gesture Recognitionno summary yetcs-cv2111.00038Google Research3 citesOct 29, 2021
- OctDIB-R++: Learning to Predict Lighting and Material with a Hybrid Differentiable Rendererno summary yetcs-cv2111.00140Google Research9 citesOct 30, 2021
- Sep4D-Net for Learned Multi-Modal Alignmentno summary yetcs-cv2109.01066Google Research0 citesSep 2, 2021
- SepRevisiting 3D ResNets for Video Recognitionno summary yetcs-cv2109.01696Google Research15 citesSep 3, 2021
- SepPR-Net: Preference Reasoning for Personalized Video Highlight Detectionno summary yetcs-cv2109.01799Tencent0 citesSep 4, 2021
- SepLearning Object-Compositional Neural Radiance Field for Editable Scene Renderingno summary yetcs-cv2109.01847Google Research6 citesSep 4, 2021
- SepParsing Table Structures in the Wildno summary yetcs-cv2109.02199Alibaba0 citesSep 6, 2021
- SepPoint-Based Neural Rendering with Per-View Optimizationno summary yetcs-cv2109.02369Adobe0 citesSep 6, 2021
- SepZero-Shot Out-of-Distribution Detection Based on the Pre-trained Model CLIPno summary yetcs-cv2109.02748Amazon102 citesSep 6, 2021
- SepPP-OCRv2: Bag of Tricks for Ultra Lightweight OCR Systemno summary yetcs-cv2109.03144Baidu26 citesSep 7, 2021
- SepOnline Unsupervised Learning of Visual Representations and Categoriesno summary yetcs-cv2109.05675Google Research0 citesSep 13, 2021
- SepLow-Shot Validation: Active Importance Sampling for Estimating Classifier Performance on Rare Categoriesno summary yetcs-cv2109.05720Google Research0 citesSep 13, 2021
- SepLearning to Ground Visual Objects for Visual Dialogno summary yetcs-cv2109.06013Google Research0 citesSep 13, 2021
- SepFrom Heatmaps to Structural Explanations of Image Classifiersno summary yetcs-cv2109.06365Tencent0 citesSep 13, 2021
- SepScalable Font Reconstruction with Dual Latent Manifoldsno summary yetcs-cv2109.06627Google Research1 citesSep 10, 2021
- SepZFlow: Gated Appearance Flow-based Virtual Try-on with 3D Priorsno summary yetcs-cv2109.07001Adobe0 citesSep 14, 2021
- SepContact-Aware Retargeting of Skinned Motionno summary yetcs-cv2109.07431Adobe1 citesSep 15, 2021
- SepNeural Human Performer: Learning Generalizable Radiance Fields for Human Performance Renderingno summary yetcs-cv2109.07448Adobe65 citesSep 15, 2021
- SepVPN: Video Provenance Network for Robust Content Attributionno summary yetcs-cv2109.10038Adobe9 citesSep 21, 2021
- SepDifferentiable Surface Triangulationno summary yetcs-cv2109.10695Adobe1 citesSep 22, 2021
- SepPairwise Emotional Relationship Recognition in Drama Videos: Dataset and Benchmarkno summary yetcs-cv2109.11243Alibaba8 citesSep 23, 2021
- SepVisually Grounded Concept Compositionno summary yetcs-cv2109.14115Google Research0 citesSep 29, 2021
- SepCrossCLR: Cross-modal Contrastive Learning For Multi-modal Video Representationsno summary yetcs-cv2109.14910Amazon0 citesSep 30, 2021
- SepPP-LCNet: A Lightweight CPU Convolutional Neural Networkno summary yetcs-cv2109.15099Baidu97 citesSep 17, 2021
- SepFake It Till You Make It: Face analysis in the wild using synthetic data aloneno summary yetcs-cv2109.15102Microsoft Research5 citesSep 30, 2021
- AugDistributed Attention for Grounded Image Captioningno summary yetcs-cv2108.01056Tencent19 citesAug 2, 2021
- AugFinding Discriminative Filters for Specific Degradations in Blind Super-Resolutionno summary yetcs-cv2108.01070Tencent19 citesAug 2, 2021
- AugConsistent Depth of Moving Objects in Videono summary yetcs-cv2108.01166Google Research26 citesAug 2, 2021
- AugVideo Similarity and Alignment Learning on Partial Video Copy Detectionno summary yetcs-cv2108.01817Alibaba2 citesAug 4, 2021
- AugEnhancing Self-supervised Video Representation Learning via Multi-level Feature Optimizationno summary yetcs-cv2108.02183Tencent0 citesAug 4, 2021
- AugTransRefer3D: Entity-and-Relation Aware Transformer for Fine-Grained 3D Visual Groundingno summary yetcs-cv2108.02388Alibaba7 citesAug 5, 2021
- AugAdaptive Normalized Representation Learning for Generalizable Face Anti-Spoofingno summary yetcs-cv2108.02667Tencent8 citesAug 5, 2021
- AugVideo Contrastive Learning with Global Contextno summary yetcs-cv2108.02722Amazon1 citesAug 5, 2021
- AugImpact of Aliasing on Generalization in Deep Convolutional Networksno summary yetcs-cv2108.03489Google Research2 citesAug 7, 2021
- AugOSCAR-Net: Object-centric Scene Graph Attention for Image Attributionno summary yetcs-cv2108.03541Adobe0 citesAug 7, 2021
- AugPaint Transformer: Feed Forward Neural Painting with Stroke Predictionno summary yetcs-cv2108.03798Baidu1 citesAug 9, 2021
- AugWeakly-Supervised Spatio-Temporal Anomaly Detection in Surveillance Videono summary yetcs-cv2108.03825Baidu7 citesAug 9, 2021
- AugRegularizing Nighttime Weirdness: Efficient Self-supervised Monocular Depth Estimation in the Darkno summary yetcs-cv2108.03830Tencent5 citesAug 9, 2021
- AugMeta Gradient Adversarial Attackno summary yetcs-cv2108.04204Tencent1 citesAug 9, 2021
- AugLearning to Cut by Watching Moviesno summary yetcs-cv2108.04294Adobe0 citesAug 9, 2021
- AugDo Datasets Have Politics? Disciplinary Values in Computer Vision Dataset Developmentno summary yetcs-cv2108.04308Google Research159 citesAug 9, 2021
- AugLearning Multi-Granular Spatio-Temporal Graph Network for Skeleton-based Action Recognitionno summary yetcs-cv2108.04536Baidu8 citesAug 10, 2021
- AugR4Dyn: Exploring Radar for Self-Supervised Monocular Depth Estimation of Dynamic Scenesno summary yetcs-cv2108.04814Google Research3 citesAug 10, 2021
- AugMetaPose: Fast 3D Pose from Multiple Views without 3D Supervisionno summary yetcs-cv2108.04869Google Research0 citesAug 10, 2021
- AugMining the Benefits of Two-stage and One-stage HOI Detectionno summary yetcs-cv2108.05077Alibaba35 citesAug 11, 2021
- AugDual Path Learning for Domain Adaptation of Semantic Segmentationno summary yetcs-cv2108.06337Microsoft Research5 citesAug 13, 2021
- AugAudio2Gestures: Generating Diverse Gestures from Speech Audio with Conditional Variational Autoencodersno summary yetcs-cv2108.06720Tencent0 citesAug 15, 2021
- AugSSH: A Self-Supervised Framework for Image Harmonizationno summary yetcs-cv2108.06805Adobe6 citesAug 15, 2021
- AugDeep Self-Adaptive Hashing for Image Retrievalno summary yetcs-cv2108.07094Tencent0 citesAug 16, 2021
- AugScene Designer: a Unified Model for Scene Search and Synthesis from Sketchno summary yetcs-cv2108.07353Adobe2 citesAug 16, 2021
- AugA New Bidirectional Unsupervised Domain Adaptation Segmentation Frameworkno summary yetcs-cv2108.07979Tencent0 citesAug 18, 2021
- AugMulti-Anchor Active Domain Adaptation for Semantic Segmentationno summary yetcs-cv2108.08012Tencent3 citesAug 18, 2021
- AugStochastic Scene-Aware Motion Predictionno summary yetcs-cv2108.08284Adobe0 citesAug 18, 2021
- AugDo Vision Transformers See Like Convolutional Neural Networks?no summary yetcs-cv2108.08810Google Research109 citesAug 19, 2021
- AugTowards Vivid and Diverse Image Colorization with Generative Color Priorno summary yetcs-cv2108.08826Tencent6 citesAug 19, 2021
- AugGraph-to-3D: End-to-End Generation and Manipulation of 3D Scenes Using Scene Graphsno summary yetcs-cv2108.08841Google Research3 citesAug 19, 2021
- AugTowards A Fairer Landmark Recognition Datasetno summary yetcs-cv2108.08874Google Research4 citesAug 19, 2021
- AugPatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibilityno summary yetcs-cv2108.08943Microsoft Research0 citesAug 19, 2021
- AugGroup-based Distinctive Image Captioning with Memory Attentionno summary yetcs-cv2108.09151Baidu1 citesAug 20, 2021
- AugEnd2End Occluded Face Recognition by Masking Corrupted Featuresno summary yetcs-cv2108.09468Tencent113 citesAug 21, 2021
- AugExploring the Quality of GAN Generated Images for Person Re-Identificationno summary yetcs-cv2108.09977Alibaba20 citesAug 23, 2021
- AugimGHUM: Implicit Generative Models of 3D Human Shape and Articulated Poseno summary yetcs-cv2108.10842Google Research4 citesAug 24, 2021
- AugGeneralize then Adapt: Source-Free Domain Adaptive Semantic Segmentationno summary yetcs-cv2108.11249Google Research2 citesAug 25, 2021
- AugSpatio-Temporal Dynamic Inference Network for Group Activity Recognitionno summary yetcs-cv2108.11743Alibaba3 citesAug 26, 2021
- AugDAE-GAN: Dynamic Aspect-aware GAN for Text-to-Image Synthesisno summary yetcs-cv2108.12141Tencent9 citesAug 27, 2021
- AugLearning Inner-Group Relations on Point Cloudsno summary yetcs-cv2108.12468Tencent0 citesAug 27, 2021
- AugNeuroCartography: Scalable Automatic Visual Summarization of Concepts in Deep Neural Networksno summary yetcs-cv2108.12931Apple1 citesAug 29, 2021
- AugEfficient Visual Recognition with Deep Neural Networks: A Survey on Recent Advances and New Directionsno summary yetcs-cv2108.13055Tencent17 citesAug 30, 2021
- AugSemIE: Semantically-aware Image Extrapolationno summary yetcs-cv2108.13702Adobe0 citesAug 31, 2021
- AugCPFN: Cascaded Primitive Fitting Networks for High-Resolution Point Cloudsno summary yetcs-cv2109.00113Adobe0 citesAug 31, 2021
- AugKnowledge Perceived Multi-modal Pretraining in E-commerceno summary yetcs-cv2109.00895Alibaba21 citesAug 20, 2021
- JulGlobal Filter Networks for Image Classificationno summary yetcs-cv2107.00645Tencent19 citesJul 1, 2021
- JulLearning Hierarchical Graph Neural Networks for Image Clusteringno summary yetcs-cv2107.01319Amazon2 citesJul 3, 2021
- JulWeb-Scale Generic Object Detection at Microsoft Bingno summary yetcs-cv2107.01814Microsoft Research0 citesJul 5, 2021
- JulContrastive Multimodal Fusion with TupleInfoNCEno summary yetcs-cv2107.02575Google Research1 citesJul 6, 2021
- JulPoseRN: A 2D pose refinement network for bias-free multi-view 3D human pose estimationno summary yetcs-cv2107.03000Microsoft Research0 citesJul 7, 2021
- JulTensor Methods in Computer Vision and Deep Learningno summary yetcs-cv2107.03436NVIDIA170 citesJul 7, 2021
- JulDeep Image Synthesis from Intuitive User Input: A Review and Perspectivesno summary yetcs-cv2107.04240Meta / FAIR1 citesJul 9, 2021
- JulMulti-Modal Association based Grouping for Form Structure Extractionno summary yetcs-cv2107.04396Adobe0 citesJul 9, 2021
- JulMultimodal Icon Annotation For Mobile Applicationsno summary yetcs-cv2107.04452Google Research0 citesJul 9, 2021
- JulViTGAN: Training GANs with Vision Transformersno summary yetcs-cv2107.04589Microsoft Research79 citesJul 9, 2021
- JulRetrieve in Style: Unsupervised Facial Feature Transfer and Retrievalno summary yetcs-cv2107.06256Google Research0 citesJul 13, 2021
- JulArtificial Intelligence in PET: an Industry Perspectiveno summary yetcs-cv2107.06747NVIDIA0 citesJul 14, 2021
- JulFace.evoLVe: A High-Performance Face Recognition Libraryno summary yetcs-cv2107.08621Baidu27 citesJul 19, 2021
- JulRECIST-Net: Lesion detection via grouping keypoints on RECIST-based annotationno summary yetcs-cv2107.08715Tencent0 citesJul 19, 2021
- JulSelf-Promoted Prototype Refinement for Few-Shot Class-Incremental Learningno summary yetcs-cv2107.08918Google Research0 citesJul 19, 2021
- JulDiscriminator-Free Generative Adversarial Attackno summary yetcs-cv2107.09225Tencent20 citesJul 20, 2021
- JulDRDF: Determining the Importance of Different Multimodal Information with Dual-Router Dynamic Frameworkno summary yetcs-cv2107.09909Alibaba0 citesJul 21, 2021
- JulCross-Sentence Temporal and Semantic Relations in Video Activity Localisationno summary yetcs-cv2107.11443Adobe2 citesJul 23, 2021
- JulSelf-Conditioned Probabilistic Learning of Video Rescalingno summary yetcs-cv2107.11639Baidu0 citesJul 24, 2021
- JulEfficient Large Scale Inlier Voting for Geometric Vision Problemsno summary yetcs-cv2107.11810Google Research0 citesJul 25, 2021
- JulHANet: Hierarchical Alignment Networks for Video-Text Retrievalno summary yetcs-cv2107.12059Alibaba8 citesJul 26, 2021
- JulLearning to Adversarially Blur Visual Object Trackingno summary yetcs-cv2107.12085Alibaba2 citesJul 26, 2021
- JulCross-modal Consensus Network for Weakly Supervised Temporal Action Localizationno summary yetcs-cv2107.12589Tencent1 citesJul 27, 2021
- JulUniformity in Heterogeneity:Diving Deep into Count Interval Partition for Crowd Countingno summary yetcs-cv2107.12619Tencent0 citesJul 27, 2021
- JulCoarse to Fine: Domain Adaptive Crowd Counting via Adversarial Scoring Networkno summary yetcs-cv2107.12858Baidu3 citesJul 27, 2021
- JulEnriching Local and Global Contexts for Temporal Action Localizationno summary yetcs-cv2107.12960Microsoft Research1 citesJul 27, 2021
- JulShape Controllable Virtual Try-on for Underwear Modelsno summary yetcs-cv2107.13156Alibaba15 citesJul 28, 2021
- JulDiscovering 3D Parts from Image Collectionsno summary yetcs-cv2107.13629Google Research0 citesJul 28, 2021
- JulVMNet: Voxel-Mesh Network for Geodesic-Aware 3D Semantic Segmentationno summary yetcs-cv2107.13824Tencent4 citesJul 29, 2021
- JulA Unified Efficient Pyramid Transformer for Semantic Segmentationno summary yetcs-cv2107.14209Amazon4 citesJul 29, 2021
- JulObject-aware Contrastive Learning for Debiased Scene Representationno summary yetcs-cv2108.00049Google Research0 citesJul 30, 2021
- JulPragmatic Image Compression for Human-in-the-Loop Decision-Makingno summary yetcs-cv2108.04219Google Research0 citesJul 7, 2021
- JunAPES: Audiovisual Person Search in Untrimmed Videono summary yetcs-cv2106.01667Adobe1 citesJun 3, 2021
- JunYou Never Cluster Aloneno summary yetcs-cv2106.01908Alibaba1 citesJun 3, 2021
- JunNeRFactor: Neural Factorization of Shape and Reflectance Under an Unknown Illuminationno summary yetcs-cv2106.01970Google Research280 citesJun 3, 2021
- JunNMS-Loss: Learning with Non-Maximum Suppression for Crowded Pedestrian Detectionno summary yetcs-cv2106.02426Tencent35 citesJun 4, 2021
- JunAligning Pretraining for Detection via Object-Level Contrastive Learningno summary yetcs-cv2106.02637Microsoft Research1 citesJun 4, 2021
- JunAssociating Objects with Transformers for Video Object Segmentationno summary yetcs-cv2106.02638Baidu19 citesJun 4, 2021
- JunGo with the Flows: Mixtures of Normalizing Flows for Point Cloud Generation and Reconstructionno summary yetcs-cv2106.03135Google Research1 citesJun 6, 2021
- JunVideo Imprintno summary yetcs-cv2106.03283Microsoft Research4 citesJun 7, 2021
- JunWide-Baseline Relative Camera Pose Estimation with Directional Learningno summary yetcs-cv2106.03336Google Research0 citesJun 7, 2021
- JunSIMONe: View-Invariant, Temporally-Abstracted Object Representations via Unsupervised Video Decompositionno summary yetcs-cv2106.03849Google Research9 citesJun 7, 2021
- JunLipSync3D: Data-Efficient Learning of Personalized 3D Talking Faces from Video using Pose and Lighting Normalizationno summary yetcs-cv2106.04185Google Research12 citesJun 8, 2021
- JunOn the relation between statistical learning and perceptual distancesno summary yetcs-cv2106.04427Google Research5 citesJun 8, 2021
- JunLow-Rank Subspaces in GANsno summary yetcs-cv2106.04488Alibaba1 citesJun 8, 2021
- JunRobustNav: Towards Benchmarking Robustness in Embodied Navigationno summary yetcs-cv2106.04531AllenAI4 citesJun 8, 2021
- JunChasing Sparsity in Vision Transformers: An End-to-End Explorationno summary yetcs-cv2106.04533Microsoft Research1 citesJun 8, 2021
- JunPAM: Understanding Product Images in Cross Product Category Attribute Extractionno summary yetcs-cv2106.04630Amazon26 citesJun 8, 2021
- JunVALUE: A Multi-Task Benchmark for Video-and-Language Understanding Evaluationno summary yetcs-cv2106.04632Microsoft Research6 citesJun 8, 2021
- JunCoAtNet: Marrying Convolution and Attention for All Data Sizesno summary yetcs-cv2106.04803Google Research1 citesJun 9, 2021
- JunSalient Object Ranking with Position-Preserved Attentionno summary yetcs-cv2106.05047Alibaba0 citesJun 9, 2021
- JunA machine learning pipeline for aiding school identification from child trafficking imagesno summary yetcs-cv2106.05215Microsoft Research1 citesJun 9, 2021
- JunImplicit-PDF: Non-Parametric Representation of Probability Distributions on the Rotation Manifoldno summary yetcs-cv2106.05965Google Research7 citesJun 10, 2021
- JunScaling Vision with Sparse Mixture of Expertsno summary yetcs-cv2106.05974Amazon29 citesJun 10, 2021
- JunSimSwap: An Efficient Framework For High Fidelity Face Swappingno summary yetcs-cv2106.06340Tencent358 citesJun 11, 2021
- JunA Multi-Implicit Neural Representation for Fontsno summary yetcs-cv2106.06866Adobe0 citesJun 12, 2021
- JunContext-Aware Image Inpainting with Learned Semantic Priorsno summary yetcs-cv2106.07220Tencent3 citesJun 14, 2021
- JunMagic Layouts: Structural Prior for Component Detection in User Interface Designsno summary yetcs-cv2106.07615Adobe0 citesJun 14, 2021
- JunImproved Transformer for High-Resolution GANsno summary yetcs-cv2106.07631Google Research3 citesJun 14, 2021
- JunLearning to Aggregate and Personalize 3D Face from In-the-Wild Photo Collectionno summary yetcs-cv2106.07852Tencent2 citesJun 15, 2021
- JunCompositional Sketch Searchno summary yetcs-cv2106.08009Adobe0 citesJun 15, 2021
- JunGradient Forward-Propagation for Large-Scale Temporal Video Modellingno summary yetcs-cv2106.08318DeepMind1 citesJun 15, 2021
- JunDynamic Head: Unifying Object Detection Heads with Attentionsno summary yetcs-cv2106.08322Microsoft Research63 citesJun 15, 2021
- JunShape from Blur: Recovering Textured 3D Shape and Motion of Fast Moving Objectsno summary yetcs-cv2106.08762Google Research1 citesJun 16, 2021
- JunOptical Mouse: 3D Mouse Pose From Single-View Videono summary yetcs-cv2106.09251Google Research0 citesJun 17, 2021
- JunTHUNDR: Transformer-based 3D HUmaN Reconstruction with Markersno summary yetcs-cv2106.09336Google Research2 citesJun 17, 2021
- JunLearning to Associate Every Segment for Video Panoptic Segmentationno summary yetcs-cv2106.09453Adobe1 citesJun 17, 2021
- JunLearning to Predict Visual Attributes in the Wildno summary yetcs-cv2106.09707Adobe0 citesJun 17, 2021
- JunGuided Integrated Gradients: An Adaptive Path Method for Removing Noiseno summary yetcs-cv2106.09788Google Research4 citesJun 17, 2021
- JunHifiFace: 3D Shape and Semantic Prior Guided High Fidelity Face Swappingno summary yetcs-cv2106.09965Tencent8 citesJun 18, 2021
- JunTowards Distraction-Robust Active Visual Trackingno summary yetcs-cv2106.10110Tencent1 citesJun 18, 2021
- JunBridging the Gap Between Object Detection and User Intent via Query-Modulationno summary yetcs-cv2106.10258Google Research0 citesJun 18, 2021
- JunEnd-to-end Temporal Action Detection with Transformerno summary yetcs-cv2106.10271Alibaba260 citesJun 18, 2021
- JunSingle View Physical Distance Estimation using Human Poseno summary yetcs-cv2106.10335Amazon0 citesJun 18, 2021
- JunOadTR: Online Action Detection with Transformersno summary yetcs-cv2106.11149Alibaba0 citesJun 21, 2021
- JunTokenLearner: What Can 8 Learned Tokens Do for Images and Videos?no summary yetcs-cv2106.11297Google Research12 citesJun 21, 2021
- JunLegoFormer: Transformers for Block-by-Block Multi-view 3D Reconstructionno summary yetcs-cv2106.12102Google Research0 citesJun 23, 2021
- JunHow Well do Feature Visualizations Support Causal Understanding of CNN Activations?no summary yetcs-cv2106.12447Amazon19 citesJun 23, 2021
- JunFusionPainting: Multimodal Fusion with Adaptive Attention for 3D Object Detectionno summary yetcs-cv2106.12449Baidu9 citesJun 23, 2021
- JunFitVid: Overfitting in Pixel-Level Video Predictionno summary yetcs-cv2106.13195Google Research29 citesJun 24, 2021
- JunMultimodal Few-Shot Learning with Frozen Language Modelsno summary yetcs-cv2106.13884Google Research86 citesJun 25, 2021
- JunRobust Pose Transfer with Dynamic Details using Neural Video Renderingno summary yetcs-cv2106.14132Tencent6 citesJun 27, 2021
- JunFeature Combination Meets Attention: Baidu Soccer Embeddings and Transformer based Temporal Detectionno summary yetcs-cv2106.14447Baidu21 citesJun 28, 2021
- JunDual Reweighting Domain Generalization for Face Presentation Attack Detectionno summary yetcs-cv2106.16128Tencent6 citesJun 30, 2021
- JunSimple Training Strategies and Model Scaling for Object Detectionno summary yetcs-cv2107.00057Google Research27 citesJun 30, 2021
- JunAttention Bottlenecks for Multimodal Fusionno summary yetcs-cv2107.00135Google Research264 citesJun 30, 2021
- MayCOMISR: Compression-Informed Video Super-Resolutionno summary yetcs-cv2105.01237Google Research0 citesMay 4, 2021
- MayInstances as Queriesno summary yetcs-cv2105.01928Tencent5 citesMay 5, 2021
- MayA Step Toward More Inclusive People Annotations for Fairnessno summary yetcs-cv2105.02317Google Research30 citesMay 5, 2021
- MayComputer-Aided Design as Languageno summary yetcs-cv2105.02769Google Research40 citesMay 6, 2021
- MayAdv-Makeup: A New Imperceptible and Transferable Attack on Face Recognitionno summary yetcs-cv2105.03162Tencent13 citesMay 7, 2021
- MayHuMoR: 3D Human Motion Model for Robust Pose Estimationno summary yetcs-cv2105.04668Adobe6 citesMay 10, 2021
- MayVICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learningno summary yetcs-cv2105.04906Meta / FAIR285 citesMay 11, 2021
- MaySegmenter: Transformer for Semantic Segmentationno summary yetcs-cv2105.05633Google Research34 citesMay 12, 2021
- MayDirectional GAN: A Novel Conditioning Strategy for Generative Networksno summary yetcs-cv2105.05712Adobe0 citesMay 12, 2021
- MayA Fast Deep Learning Network for Automatic Image Auto-Straighteningno summary yetcs-cv2105.05787Adobe0 citesMay 12, 2021
- MaySemantic Diversity Learning for Zero-Shot Multi-label Classificationno summary yetcs-cv2105.05926Alibaba0 citesMay 12, 2021
- MayCompatibility-aware Heterogeneous Visual Searchno summary yetcs-cv2105.06047Amazon0 citesMay 13, 2021
- MayDeep Unsupervised Hashing by Distilled Smooth Guidanceno summary yetcs-cv2105.06125Alibaba0 citesMay 13, 2021
- MayEpisodic Transformer for Vision-and-Language Navigationno summary yetcs-cv2105.06453Google Research9 citesMay 13, 2021
- MayAutomatic Non-Linear Video Editing Transferno summary yetcs-cv2105.06988Google Research0 citesMay 14, 2021
- MaySMURF: Self-Teaching Multi-Frame Unsupervised RAFT with Full-Image Warpingno summary yetcs-cv2105.07014Google Research0 citesMay 14, 2021
- MayAre Convolutional Neural Networks or Transformers more like human vision?no summary yetcs-cv2105.07197Google Research21 citesMay 15, 2021
- MayVisual FUDGE: Form Understanding via Dynamic Graph Editingno summary yetcs-cv2105.08194Adobe1 citesMay 17, 2021
- MayWeakly Supervised Dense Video Captioning via Jointly Usage of Knowledge Distillation and Cross-modal Matchingno summary yetcs-cv2105.08252Baidu1 citesMay 18, 2021
- MayPathdreamer: A World Model for Indoor Navigationno summary yetcs-cv2105.08756Google Research3 citesMay 18, 2021
- MayJoint Face Image Restoration and Frontalization for Recognitionno summary yetcs-cv2105.09907Baidu71 citesMay 12, 2021
- MayExploring Robustness of Unsupervised Domain Adaptation in Semantic Segmentationno summary yetcs-cv2105.10843Tencent0 citesMay 23, 2021
- MaySiamRCR: Reciprocal Classification and Regression for Visual Object Trackingno summary yetcs-cv2105.11237Tencent2 citesMay 24, 2021
- MayLineCounter: Learning Handwritten Text Line Segmentation by Countingno summary yetcs-cv2105.11307Amazon0 citesMay 24, 2021
- MayAttention-guided Temporally Coherent Video Object Mattingno summary yetcs-cv2105.11427Alibaba28 citesMay 24, 2021
- MayFILTRA: Rethinking Steerable CNN by Filter Transformno summary yetcs-cv2105.11636Baidu0 citesMay 25, 2021
- MayStyle Similarity as Feedback for Product Designno summary yetcs-cv2105.12256Meta / FAIR1 citesMay 25, 2021
- MayKLIEP-based Density Ratio Estimation for Semantically Consistent Synthetic to Real Images Adaptation in Urban Traffic Scenesno summary yetcs-cv2105.12549Google Research1 citesMay 26, 2021
- MayBlind Motion Deblurring Super-Resolution: When Dynamic Spatio-Temporal Learning Meets Static Image Understandingno summary yetcs-cv2105.13077Tencent22 citesMay 27, 2021
- MayImproving Facial Attribute Recognition by Group and Graph Learningno summary yetcs-cv2105.13825Tencent0 citesMay 28, 2021
- MayBoosting Monocular Depth Estimation Models to High-Resolution via Content-Adaptive Multi-Resolution Mergingno summary yetcs-cv2105.14021Adobe29 citesMay 28, 2021
- MayM6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformersno summary yetcs-cv2105.14211Alibaba23 citesMay 29, 2021
- MayPolygonal Point Set Trackingno summary yetcs-cv2105.14584Adobe0 citesMay 30, 2021
- MayDual-stream Network for Visual Recognitionno summary yetcs-cv2105.14734Baidu0 citesMay 31, 2021
- MayAnalogous to Evolutionary Algorithm: Designing a Unified Sequence Modelno summary yetcs-cv2105.15089Tencent13 citesMay 31, 2021
- MaySimilarity Embedding Networks for Robust Human Activity Recognitionno summary yetcs-cv2106.15283Tencent11 citesMay 31, 2021
- AprSelf-supervised Motion Learning from Static Imagesno summary yetcs-cv2104.00240Alibaba7 citesApr 1, 2021
- AprFrozen in Time: A Joint Video and Image Encoder for End-to-End Retrievalno summary yetcs-cv2104.00650Google Research24 citesApr 1, 2021
- AprMemorability: An image-computable measure of information utilityno summary yetcs-cv2104.00805Adobe4 citesApr 1, 2021
- AprDefending Against Image Corruptions Through Adversarial Augmentationsno summary yetcs-cv2104.01086Google Research4 citesApr 2, 2021
- AprGraph Contrastive Clusteringno summary yetcs-cv2104.01429Alibaba0 citesApr 3, 2021
- AprContent-Aware GAN Compressionno summary yetcs-cv2104.02244Adobe2 citesApr 6, 2021
- AprScene Graph Embeddings Using Relative Similarity Supervisionno summary yetcs-cv2104.02381Adobe1 citesApr 6, 2021
- AprVariational Transformer Networks for Layout Generationno summary yetcs-cv2104.02416Google Research8 citesApr 6, 2021
- AprFine-Grained Fashion Similarity Prediction by Attribute-Specific Embedding Learningno summary yetcs-cv2104.02429Alibaba44 citesApr 6, 2021
- AprgradSim: Differentiable simulation for system identification and visuomotor controlno summary yetcs-cv2104.02646Google Research44 citesApr 6, 2021
- AprDifferentiable Patch Selection for Image Recognitionno summary yetcs-cv2104.03059Google Research9 citesApr 7, 2021
- AprDoes Your Dermatology Classifier Know What It Doesn't Know? Detecting the Long-Tail of Unseen Conditionsno summary yetcs-cv2104.03829DeepMind16 citesApr 8, 2021
- AprField Convolutions for Surface CNNsno summary yetcs-cv2104.03916Adobe2 citesApr 8, 2021
- AprDe-rendering the World's Revolutionary Artefactsno summary yetcs-cv2104.03954Google Research0 citesApr 8, 2021
- AprModulated Periodic Activations for Generalizable Local Functional Representationsno summary yetcs-cv2104.03960Adobe1 citesApr 8, 2021
- AprCutPaste: Self-Supervised Learning for Anomaly Detection and Localizationno summary yetcs-cv2104.04015Google Research87 citesApr 8, 2021
- AprSpatially-Varying Outdoor Lighting Estimation from Intrinsicsno summary yetcs-cv2104.04160Google Research0 citesApr 9, 2021
- AprA Reinforcement-Learning-Based Energy-Efficient Framework for Multi-Task Video Analytics Pipelineno summary yetcs-cv2104.04443Alibaba13 citesApr 9, 2021
- AprEgocentric Pose Estimation from Human Vision Spanno summary yetcs-cv2104.05167Meta / FAIR0 citesApr 12, 2021
- AprDrafting and Revision: Laplacian Pyramid Network for Fast High-Quality Artistic Style Transferno summary yetcs-cv2104.05376Baidu10 citesApr 12, 2021
- AprTowards Efficient Graph Convolutional Networks for Point Cloud Handlingno summary yetcs-cv2104.05706Microsoft Research0 citesApr 12, 2021
- AprIMAGINE: Image Synthesis by Image-Guided Model Inversionno summary yetcs-cv2104.05895Adobe5 citesApr 13, 2021
- AprVariTex: Variational Neural Face Texturesno summary yetcs-cv2104.05988Google Research0 citesApr 13, 2021
- AprZeus: Efficiently Localizing Actions in Videos using Reinforcement Learningno summary yetcs-cv2104.06142Microsoft Research9 citesApr 6, 2021
- AprLearning Semantic Person Image Generation by Region-Adaptive Normalizationno summary yetcs-cv2104.06650Baidu4 citesApr 14, 2021
- AprRevisiting Hierarchical Approach for Persistent Long-Term Video Predictionno summary yetcs-cv2104.06697Google Research13 citesApr 14, 2021
- AprTemporally-Coherent Surface Reconstruction via Metric-Consistent Atlasesno summary yetcs-cv2104.06950Adobe0 citesApr 14, 2021
- AprAdaptive Intermediate Representations for Video Understandingno summary yetcs-cv2104.07135Google Research0 citesApr 14, 2021
- AprE2Style: Improve the Efficiency and Effectiveness of StyleGAN Inversionno summary yetcs-cv2104.07661Microsoft Research62 citesApr 15, 2021
- AprRethinking Text Line Recognition Modelsno summary yetcs-cv2104.07787Google Research19 citesApr 15, 2021
- AprVGNMN: Video-grounded Neural Module Network to Video-Grounded Language Tasksno summary yetcs-cv2104.07921Salesforce0 citesApr 16, 2021
- AprData-Driven 3D Reconstruction of Dressed Humans From Sparse Viewsno summary yetcs-cv2104.08013Meta / FAIR17 citesApr 16, 2021
- AprRPCL: A Framework for Improving Cross-Domain Detection with Auxiliary Tasksno summary yetcs-cv2104.08689Adobe1 citesApr 18, 2021
- AprECACL: A Holistic Framework for Semi-Supervised Domain Adaptationno summary yetcs-cv2104.09136Adobe4 citesApr 19, 2021
- AprLarge Scale Interactive Motion Forecasting for Autonomous Driving : The Waymo Open Motion Datasetno summary yetcs-cv2104.10133Google Research1 citesApr 20, 2021
- AprImproving Weakly-supervised Object Localization via Causal Interventionno summary yetcs-cv2104.10351Baidu0 citesApr 21, 2021
- AprTowards Adversarial Patch Analysis and Certified Defense against Crowd Countingno summary yetcs-cv2104.10868Baidu2 citesApr 22, 2021
- AprVATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Textno summary yetcs-cv2104.11178Google Research340 citesApr 22, 2021
- AprH2O: Two Hands Manipulating Objects for First Person Interaction Recognitionno summary yetcs-cv2104.11181Microsoft Research6 citesApr 22, 2021
- AprKeypointDeformer: Unsupervised 3D Keypoint Discovery for Shape Controlno summary yetcs-cv2104.11224Google Research3 citesApr 22, 2021
- AprVector Neurons: A General Framework for SO(3)-Equivariant Networksno summary yetcs-cv2104.12229Google Research3 citesApr 25, 2021
- AprVCGAN: Video Colorization with Hybrid Generative Adversarial Networkno summary yetcs-cv2104.12357Tencent40 citesApr 26, 2021
- AprDelving into Data: Effectively Substitute Training for Black-box Attackno summary yetcs-cv2104.12378Tencent3 citesApr 26, 2021
- Apr3D Scene Compression through Entropy Penalized Neural Representation Functionsno summary yetcs-cv2104.12456Google Research1 citesApr 26, 2021
- AprLess is more: Selecting informative and diverse subsets with balancing constraintsno summary yetcs-cv2104.12835Google Research2 citesApr 26, 2021
- AprNTIRE 2021 Depth Guided Image Relighting Challengeno summary yetcs-cv2104.13365Baidu9 citesApr 27, 2021
- AprPAFNet: An Efficient Anchor-Free Object Detector Guidanceno summary yetcs-cv2104.13534Baidu6 citesApr 28, 2021
- AprImage Inpainting by End-to-End Cascaded Refinement with Mask Awarenessno summary yetcs-cv2104.13743Baidu133 citesApr 28, 2021
- AprWith a Little Help from My Friends: Nearest-Neighbor Contrastive Learning of Visual Representationsno summary yetcs-cv2104.14548Google Research3 citesApr 29, 2021
- AprMarioNette: Self-Supervised Sprite Learningno summary yetcs-cv2104.14553Adobe0 citesApr 29, 2021
- AprLearning Multi-Granular Hypergraphs for Video-Based Person Re-Identificationno summary yetcs-cv2104.14913Tencent1 citesApr 30, 2021
- AprDifferentiable Event Stream Simulator for Non-Rigid 3D Trackingno summary yetcs-cv2104.15139Google Research0 citesApr 30, 2021
- MarRepresentation Learning for Event-based Visuomotor Policiesno summary yetcs-cv2103.00806Microsoft Research0 citesMar 1, 2021
- MarOmniNet: Omnidirectional Representations from Transformersno summary yetcs-cv2103.01075Google Research9 citesMar 1, 2021
- MarDeep Perceptual Image Quality Assessment for Compressionno summary yetcs-cv2103.01114Google Research2 citesMar 1, 2021
- MarA Deep Emulator for Secondary Motion of 3D Charactersno summary yetcs-cv2103.01261Adobe1 citesMar 1, 2021
- MarScalable Scene Flow from Point Clouds in the Real Worldno summary yetcs-cv2103.01306Google Research2 citesMar 1, 2021
- MarDepth from Camera Motion and Object Detectionno summary yetcs-cv2103.01468AllenAI1 citesMar 2, 2021
- MarAsk&Confirm: Active Detail Enriching for Cross-Modal Retrieval with Partial Queryno summary yetcs-cv2103.01654Tencent1 citesMar 2, 2021
- MarPredicting Video with VQVAEno summary yetcs-cv2103.01950DeepMind26 citesMar 2, 2021
- MarAdaptive Consistency Regularization for Semi-Supervised Transfer Learningno summary yetcs-cv2103.02193Baidu8 citesMar 3, 2021
- MarDeepFN: Towards Generalizable Facial Action Unit Recognition with Deep Face Normalizationno summary yetcs-cv2103.02484Microsoft Research4 citesMar 3, 2021
- MarDONeRF: Towards Real-Time Rendering of Compact Neural Radiance Fields using Depth Oracle Networksno summary yetcs-cv2103.03231Meta / FAIR248 citesMar 4, 2021
- MarNutrition5k: Towards Automatic Nutritional Understanding of Generic Foodno summary yetcs-cv2103.03375Google Research12 citesMar 4, 2021
- MarMeasuring Model Biases in the Absence of Ground Truthno summary yetcs-cv2103.03417Google Research15 citesMar 5, 2021
- MarStructured Scene Memory for Vision-Language Navigationno summary yetcs-cv2103.03454Salesforce12 citesMar 5, 2021
- MarGenerating Images with Sparse Representationsno summary yetcs-cv2103.03841DeepMind9 citesMar 5, 2021
- MarHigh Perceptual Quality Image Denoising with a Posterior Sampling CGANno summary yetcs-cv2103.04192Google Research5 citesMar 6, 2021
- MarDeep Model Intellectual Property Protection via Deep Watermarkingno summary yetcs-cv2103.04980Microsoft Research0 citesMar 8, 2021
- MarStabilized Medical Image Attacksno summary yetcs-cv2103.05232Tencent16 citesMar 9, 2021
- MarVideoMoCo: Contrastive Video Representation Learning with Temporally Adversarial Examplesno summary yetcs-cv2103.05905Tencent22 citesMar 10, 2021
- MarES-Net: Erasing Salient Parts to Learn More in Re-Identificationno summary yetcs-cv2103.05918Alibaba28 citesMar 10, 2021
- MarHolistic 3D Scene Understanding from a Single Image with Implicit Representationno summary yetcs-cv2103.06422Google Research2 citesMar 11, 2021
- MarDual Attention-in-Attention Model for Joint Rain Streak and Raindrop Removalno summary yetcs-cv2103.07051Tencent105 citesMar 12, 2021
- MarBoundary Proposal Network for Two-Stage Natural Language Video Localizationno summary yetcs-cv2103.08109Tencent17 citesMar 15, 2021
- MarLARNet: Lie Algebra Residual Network for Face Recognitionno summary yetcs-cv2103.08147Tencent24 citesMar 15, 2021
- MarDisentangled Cycle Consistency for Highly-realistic Virtual Try-Onno summary yetcs-cv2103.09479Tencent4 citesMar 17, 2021
- MarLearning to Resize Images for Computer Vision Tasksno summary yetcs-cv2103.09950Google Research9 citesMar 17, 2021
- MarFastNeRF: High-Fidelity Neural Rendering at 200FPSno summary yetcs-cv2103.10380Microsoft Research42 citesMar 18, 2021
- MarEfficient Visual Pretraining with Contrastive Detectionno summary yetcs-cv2103.10957DeepMind33 citesMar 19, 2021
- MarConditional Training with Bounding Map for Universal Lesion Detectionno summary yetcs-cv2103.12277Alibaba1 citesMar 23, 2021
- MarPanGEA: The Panoramic Graph Environment Annotation Toolkitno summary yetcs-cv2103.12703Google Research1 citesMar 23, 2021
- MarScaling Local Self-Attention for Parameter Efficient Visual Backbonesno summary yetcs-cv2103.12731Google Research35 citesMar 23, 2021
- MarFactors of Influence for Transfer Learning across Diverse Appearance Domains and Task Typesno summary yetcs-cv2103.13318Google Research14 citesMar 24, 2021
- MarMip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance Fieldsno summary yetcs-cv2103.13415Google Research90 citesMar 24, 2021
- MarMBA-VO: Motion Blur Aware Visual Odometryno summary yetcs-cv2103.13684Microsoft Research0 citesMar 25, 2021
- MarPatch Craft: Video Denoising by Deep Modeling and Patch Matchingno summary yetcs-cv2103.13767Google Research0 citesMar 25, 2021
- MarStepwise Goal-Driven Networks for Trajectory Predictionno summary yetcs-cv2103.14107Amazon13 citesMar 25, 2021
- MarCOTR: Correspondence Transformer for Matching Across Imagesno summary yetcs-cv2103.14167Google Research8 citesMar 25, 2021
- MarUnderstanding Robustness of Transformers for Image Classificationno summary yetcs-cv2103.14586Google Research2 citesMar 26, 2021
- MarBaking Neural Radiance Fields for Real-Time View Synthesisno summary yetcs-cv2103.14645Google Research8 citesMar 26, 2021
- MarTS-CAM: Token Semantic Coupled Attention Map for Weakly Supervised Object Localizationno summary yetcs-cv2103.14862Tencent2 citesMar 27, 2021
- MarSceneGraphFusion: Incremental 3D Scene Graph Prediction from RGB-D Sequencesno summary yetcs-cv2103.14898Google Research8 citesMar 27, 2021
- MarLabels4Free: Unsupervised Segmentation using StyleGANno summary yetcs-cv2103.14968Adobe6 citesMar 27, 2021
- MarNoise Injection-based Regularization for Point Cloud Processingno summary yetcs-cv2103.15027Amazon1 citesMar 28, 2021
- MarManhattanSLAM: Robust Planar Tracking and Mapping Leveraging Mixture of Manhattan Framesno summary yetcs-cv2103.15068Google Research6 citesMar 28, 2021
- MarMining Latent Classes for Few-shot Segmentationno summary yetcs-cv2103.15402Tencent4 citesMar 29, 2021
- MarUnified Graph Structured Models for Video Understandingno summary yetcs-cv2103.15662Google Research0 citesMar 29, 2021
- MarViViT: A Video Vision Transformerno summary yetcs-cv2103.15691Google Research79 citesMar 29, 2021
- MarTransFill: Reference-guided Image Inpainting by Merging Multiple Color and Spatial Transformationsno summary yetcs-cv2103.15982Adobe4 citesMar 29, 2021
- MarLarge Scale Autonomous Driving Scenarios Clustering with Self-supervised Feature Extractionno summary yetcs-cv2103.16101Baidu0 citesMar 30, 2021
- MarGeneralized Organ Segmentation by Imitating One-shot Reasoning using Anatomical Correlationno summary yetcs-cv2103.16344Tencent0 citesMar 30, 2021
- MarBenchmarking Representation Learning for Natural World Image Collectionsno summary yetcs-cv2103.16483Google Research14 citesMar 30, 2021
- MarThe Elastic Lottery Ticket Hypothesisno summary yetcs-cv2103.16547Microsoft Research1 citesMar 30, 2021
- MarBroaden Your Views for Self-Supervised Video Learningno summary yetcs-cv2103.16559DeepMind13 citesMar 30, 2021
- MarNeural Surface Mapsno summary yetcs-cv2103.16942Adobe1 citesMar 31, 2021
- MarStyleCLIP: Text-Driven Manipulation of StyleGAN Imageryno summary yetcs-cv2103.17249Adobe87 citesMar 31, 2021
- FebUnsupervised Novel View Synthesis from a Single Imageno summary yetcs-cv2102.03285Google Research2 citesFeb 5, 2021
- FebUniFuse: Unidirectional Fusion for 360$^{\circ}$ Panorama Depth Estimationno summary yetcs-cv2102.03550Alibaba111 citesFeb 6, 2021
- FebColorization Transformerno summary yetcs-cv2102.04432Google Research5 citesFeb 8, 2021
- FebLarge Scale Long-tailed Product Recognition System at Alibabano summary yetcs-cv2102.04652Alibaba4 citesFeb 9, 2021
- FebVirtual ID Discovery from E-commerce Media at Alibaba: Exploiting Richness of User Click Behavior for Visual Search Relevanceno summary yetcs-cv2102.04667Alibaba4 citesFeb 9, 2021
- FebVisual Search at Alibabano summary yetcs-cv2102.04674Alibaba99 citesFeb 9, 2021
- FebTelling the What while Pointing to the Where: Multimodal Queries for Image Retrievalno summary yetcs-cv2102.04980Google Research1 citesFeb 9, 2021
- FebDetecting Localized Adversarial Examples: A Generic Approach using Critical Region Analysisno summary yetcs-cv2102.05241Baidu5 citesFeb 10, 2021
- FebHigh-Performance Large-Scale Image Recognition Without Normalizationno summary yetcs-cv2102.06171DeepMind256 citesFeb 11, 2021
- FebHybrid Neural Fusion for Full-frame Video Stabilizationno summary yetcs-cv2102.06205Google Research1 citesFeb 11, 2021
- FebJust Noticeable Difference for Deep Machine Visionno summary yetcs-cv2102.08168Alibaba34 citesFeb 16, 2021
- FebLambdaNetworks: Modeling Long-Range Interactions Without Attentionno summary yetcs-cv2102.08602Google Research48 citesFeb 17, 2021
- FebShaRF: Shape-conditioned Radiance Fields from a Single Viewno summary yetcs-cv2102.08860Google Research5 citesFeb 17, 2021
- FebMosaicOS: A Simple and Effective Use of Object-Centric Images for Long-Tailed Object Detectionno summary yetcs-cv2102.08884Google Research2 citesFeb 17, 2021
- FebMobile Computational Photography: A Tourno summary yetcs-cv2102.09000Google Research3 citesFeb 17, 2021
- FebCReST: A Class-Rebalancing Self-Training Framework for Imbalanced Semi-Supervised Learningno summary yetcs-cv2102.09559Google Research33 citesFeb 18, 2021
- FebCamera Calibration with Pose Guidanceno summary yetcs-cv2102.10202NVIDIA0 citesFeb 19, 2021
- FebGMLight: Lighting Estimation via Geometric Distribution Approximationno summary yetcs-cv2102.10244Alibaba26 citesFeb 20, 2021
- FebTowards Accurate and Compact Architectures via Neural Architecture Transformerno summary yetcs-cv2102.10301Tencent1 citesFeb 20, 2021
- FebDecoupled and Memory-Reinforced Networks: Towards Effective Feature Learning for One-Step Person Searchno summary yetcs-cv2102.10795Baidu5 citesFeb 22, 2021
- FebSTEP: Segmenting and Tracking Every Pixelno summary yetcs-cv2102.11859Google Research9 citesFeb 23, 2021
- FebWalk2Map: Extracting Floor Plans from Indoor Walk Trajectoriesno summary yetcs-cv2103.00262Adobe0 citesFeb 27, 2021
- JanSpotPatch: Parameter-Efficient Transfer Learning for Mobile Object Detectionno summary yetcs-cv2101.01260Google Research0 citesJan 4, 2021
- JanDeep Class-Specific Affinity-Guided Convolutional Network for Multimodal Unpaired Image Segmentationno summary yetcs-cv2101.01513NVIDIA0 citesJan 5, 2021
- JanMulti-Stage Residual Hiding for Image-into-Audio Steganographyno summary yetcs-cv2101.01872Alibaba1 citesJan 6, 2021
- JanTryOnGAN: Body-Aware Try-On via Layered Interpolationno summary yetcs-cv2101.02285Google Research4 citesJan 6, 2021
- JanDiminishing Uncertainty within the Training Pool: Active Learning for Medical Image Segmentationno summary yetcs-cv2101.02323NVIDIA106 citesJan 7, 2021
- JanLearning Temporal Dynamics from Cycles in Narrated Videono summary yetcs-cv2101.02337Google Research0 citesJan 7, 2021
- JanGAN-Control: Explicitly Controllable GANsno summary yetcs-cv2101.02477Amazon12 citesJan 7, 2021
- JanMAAS: Multi-modal Assignation for Active Speaker Detectionno summary yetcs-cv2101.03682Adobe9 citesJan 11, 2021
- JanHorizontal-to-Vertical Video Conversionno summary yetcs-cv2101.04051Alibaba1 citesJan 11, 2021
- JanPaddleSeg: A High-Efficient Development Toolkit for Image Segmentationno summary yetcs-cv2101.06175Baidu60 citesJan 15, 2021
- JanCross-modal Learning for Domain Adaptation in 3D Semantic Segmentationno summary yetcs-cv2101.07253Amazon66 citesJan 18, 2021
- JanSOSD-Net: Joint Semantic Object Segmentation and Depth Estimation from Monocular imagesno summary yetcs-cv2101.07422Baidu3 citesJan 19, 2021
- JanJoint Learning of 3D Shape Retrieval and Deformationno summary yetcs-cv2101.07889Adobe5 citesJan 19, 2021
- JanAI Choreographer: Music Conditioned 3D Dance Generation with AIST++no summary yetcs-cv2101.08779Google Research32 citesJan 21, 2021
- JanThe Role of Edges in Line Drawing Perceptionno summary yetcs-cv2101.09376Adobe13 citesJan 22, 2021
- JanSupervision by Registration and Triangulation for Landmark Detectionno summary yetcs-cv2101.09866Meta / FAIR46 citesJan 25, 2021
- JanAdversarial Text-to-Image Synthesis: A Reviewno summary yetcs-cv2101.09983Adobe202 citesJan 25, 2021
- JanHexCNN: A Framework for Native Hexagonal Convolutional Neural Networksno summary yetcs-cv2101.10897Google Research0 citesJan 25, 2021
- JanRAPIQUE: Rapid and Accurate Video Quality Prediction of User Generated Contentno summary yetcs-cv2101.10955Google Research5 citesJan 26, 2021
- JanMulti-Instance Pose Networks: Rethinking Top-Down Pose Estimationno summary yetcs-cv2101.11223Amazon8 citesJan 27, 2021
- JanBottleneck Transformers for Visual Recognitionno summary yetcs-cv2101.11605Google Research33 citesJan 27, 2021
- JanMulti-Modal Aesthetic Assessment for MObile Gaming Imageno summary yetcs-cv2101.11700Tencent6 citesJan 27, 2021
- JanObject Detection Made Simpler by Eliminating Heuristic NMSno summary yetcs-cv2101.11782Alibaba39 citesJan 28, 2021
- JanComplementary Pseudo Labels For Unsupervised Domain Adaptation On Person Re-identificationno summary yetcs-cv2101.12521Alibaba84 citesJan 29, 2021
2020
573- DecJust Ask: Learning to Answer Questions from Millions of Narrated Videosno summary yetcs-cv2012.00451DeepMind14 citesDec 1, 2020
- DecDeFMO: Deblurring and Shape Recovery of Fast Moving Objectsno summary yetcs-cv2012.00595Microsoft Research5 citesDec 1, 2020
- DecGLEAN: Generative Latent Bank for Large-Factor Image Super-Resolutionno summary yetcs-cv2012.00739Tencent15 citesDec 1, 2020
- DecLearning Delaunay Surface Elements for Mesh Reconstructionno summary yetcs-cv2012.01203Adobe2 citesDec 2, 2020
- DecLearning Spatial Attention for Face Super-Resolutionno summary yetcs-cv2012.01211Tencent208 citesDec 2, 2020
- DecCross-Descriptor Visual Localization and Mappingno summary yetcs-cv2012.01377Microsoft Research1 citesDec 2, 2020
- DecWisdom of Committees: An Overlooked Approach To Faster and More Accurate Modelsno summary yetcs-cv2012.01988Google Research21 citesDec 3, 2020
- DecDeepVideoMVS: Multi-View Stereo on Video with Recurrent Spatio-Temporal Fusionno summary yetcs-cv2012.02177Microsoft Research10 citesDec 3, 2020
- DecBasicVSR: The Search for Essential Components in Video Super-Resolution and Beyondno summary yetcs-cv2012.02181Tencent32 citesDec 3, 2020
- DecEVRNet: Efficient Video Restoration on Edge Devicesno summary yetcs-cv2012.02228Meta / FAIR2 citesDec 3, 2020
- DecUnderstanding Guided Image Captioning Performance across Domainsno summary yetcs-cv2012.02339Google Research0 citesDec 4, 2020
- DecSpatial-Temporal Alignment Network for Action Recognition and Detectionno summary yetcs-cv2012.02426Google Research3 citesDec 4, 2020
- DecPractical No-box Adversarial Attacks against DNNsno summary yetcs-cv2012.02525Microsoft Research7 citesDec 4, 2020
- DecFew-shot Image Generation with Elastic Weight Consolidationno summary yetcs-cv2012.02780Adobe26 citesDec 4, 2020
- DecSpatially-Adaptive Pixelwise Networks for Fast Image Translationno summary yetcs-cv2012.02992Adobe9 citesDec 5, 2020
- DecNeRD: Neural Reflectance Decomposition from Image Collectionsno summary yetcs-cv2012.03918Google Research28 citesDec 7, 2020
- DecNeRV: Neural Reflectance and Visibility Fields for Relighting and View Synthesisno summary yetcs-cv2012.03927Google Research33 citesDec 7, 2020
- DecScale Aware Adaptation for Land-Cover Classification in Remote Sensing Imageryno summary yetcs-cv2012.04222Amazon1 citesDec 8, 2020
- DecData InStance Prior (DISP) in Generative Adversarial Networksno summary yetcs-cv2012.04256Adobe0 citesDec 8, 2020
- DecCASTing Your Model: Learning to Localize Improves Self-Supervised Representationsno summary yetcs-cv2012.04630Salesforce13 citesDec 8, 2020
- DecLook Before you Speak: Visually Contextualized Utterancesno summary yetcs-cv2012.05710Google Research8 citesDec 10, 2020
- DecLayoutGMN: Neural Graph Matching for Structural Layout Similarityno summary yetcs-cv2012.06547Adobe4 citesDec 11, 2020
- DecUncalibrated Neural Inverse Rendering for Photometric Stereo of General Surfacesno summary yetcs-cv2012.06777Google Research2 citesDec 12, 2020
- DecGeoNet++: Iterative Geometric Neural Network with Edge-Aware Refinement for Joint Depth and Surface Normal Estimationno summary yetcs-cv2012.06980Tencent64 citesDec 13, 2020
- DecSimple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentationno summary yetcs-cv2012.07177Google Research88 citesDec 13, 2020
- DecSeeing Behind Objects for 3D Multi-Object Tracking in RGB-D Sequencesno summary yetcs-cv2012.08197Adobe1 citesDec 15, 2020
- DecFMODetect: Robust Detection of Fast Moving Objectsno summary yetcs-cv2012.08216Microsoft Research1 citesDec 15, 2020
- DecAttention over learned object embeddings enables complex visual reasoningno summary yetcs-cv2012.08508Google Research7 citesDec 15, 2020
- DecProjected Distribution Loss for Image Enhancementno summary yetcs-cv2012.09289Google Research2 citesDec 16, 2020
- DecPolyblur: Removing mild blur by polynomial reblurringno summary yetcs-cv2012.09322Google Research3 citesDec 16, 2020
- DecEmbodied Visual Active Learning for Semantic Segmentationno summary yetcs-cv2012.09503Google Research1 citesDec 17, 2020
- DecObjectron: A Large Scale Dataset of Object-Centric Videos in the Wild with Pose Annotationsno summary yetcs-cv2012.09988Google Research11 citesDec 18, 2020
- Dec3D Object Detection with Pointformerno summary yetcs-cv2012.11409Amazon9 citesDec 21, 2020
- DecFrom Points to Multi-Object 3D Reconstructionno summary yetcs-cv2012.11575Google Research0 citesDec 21, 2020
- DecHuman Action Recognition from Various Data Modalities: A Reviewno summary yetcs-cv2012.11866Alibaba563 citesDec 22, 2020
- DecTime-Travel Rephotographyno summary yetcs-cv2012.12261Adobe1 citesDec 22, 2020
- DecEfficient video annotation with visual interpolation and frame selection guidanceno summary yetcs-cv2012.12554Google Research1 citesDec 23, 2020
- DecPhysics-based Shadow Image Decomposition for Shadow Removalno summary yetcs-cv2012.13018Amazon9 citesDec 23, 2020
- DecDeepSurfels: Learning Online Appearance Fusionno summary yetcs-cv2012.14240Microsoft Research0 citesDec 28, 2020
- DecLearned Multi-Resolution Variable-Rate Image Compression with Octave-based Residual Blocksno summary yetcs-cv2012.15463Google Research0 citesDec 31, 2020
- NovRevisiting Adaptive Convolutions for Video Frame Interpolationno summary yetcs-cv2011.01280Adobe2 citesNov 2, 2020
- NovSelfPose: 3D Egocentric Pose Estimation from a Headset Mounted Camerano summary yetcs-cv2011.01519Meta / FAIR75 citesNov 2, 2020
- NovDeep Image Compositingno summary yetcs-cv2011.02146Adobe0 citesNov 4, 2020
- NovPixel-wise Dense Detector for Image Inpaintingno summary yetcs-cv2011.02293Tencent0 citesNov 4, 2020
- NovLearning and Evaluating Representations for Deep One-class Classificationno summary yetcs-cv2011.02578Google Research92 citesNov 4, 2020
- NovText-to-Image Generation Grounded by Fine-Grained User Attentionno summary yetcs-cv2011.03775Google Research3 citesNov 7, 2020
- NovAdaptive Linear Span Network for Object Skeleton Detectionno summary yetcs-cv2011.03972Alibaba23 citesNov 8, 2020
- NovLearning to Infer Semantic Parameters for 3D Shape Editingno summary yetcs-cv2011.04755Google Research1 citesNov 9, 2020
- NovDebugging Tests for Model Explanationsno summary yetcs-cv2011.05429Google Research7 citesNov 10, 2020
- NovDomain Adaptation Gaze Estimation by Embedding with Prediction Consistencyno summary yetcs-cv2011.07526Tencent7 citesNov 15, 2020
- NovA Divide et Impera Approach for 3D Shape Reconstruction from Multiple Viewsno summary yetcs-cv2011.08534Google Research0 citesNov 17, 2020
- NovMultimodal Prototypical Networks for Few-shot Learningno summary yetcs-cv2011.08899Amazon7 citesNov 17, 2020
- NovLiquid Warping GAN with Attention: A Unified Framework for Human Image Synthesisno summary yetcs-cv2011.09055Tencent1 citesNov 18, 2020
- NovPositive-Congruent Training: Towards Regression-Free Model Updatesno summary yetcs-cv2011.09161Amazon5 citesNov 18, 2020
- NovModeling Fashion Influence from Photosno summary yetcs-cv2011.09663Meta / FAIR0 citesNov 17, 2020
- NovAn Effective Anti-Aliasing Approach for Residual Networksno summary yetcs-cv2011.10675Google Research23 citesNov 20, 2020
- NovRank-smoothed Pairwise Learning In Perceptual Quality Assessmentno summary yetcs-cv2011.10893Google Research0 citesNov 21, 2020
- NovLearnable Boundary Guided Adversarial Trainingno summary yetcs-cv2011.11164Tencent4 citesNov 23, 2020
- NovAdversarial Refinement Network for Human Motion Predictionno summary yetcs-cv2011.11221Tencent0 citesNov 23, 2020
- NovRobust image stitching with multiple registrationsno summary yetcs-cv2011.11784Google Research0 citesNov 23, 2020
- NovObject-centered image stitchingno summary yetcs-cv2011.11789Google Research0 citesNov 23, 2020
- NovCross-Camera Convolutional Color Constancyno summary yetcs-cv2011.11890Google Research1 citesNov 24, 2020
- NovTemporal Action Detection with Multi-level Supervisionno summary yetcs-cv2011.11893Microsoft Research1 citesNov 24, 2020
- NovDeRF: Decomposed Radiance Fieldsno summary yetcs-cv2011.12490Google Research3 citesNov 25, 2020
- NovStyleSpace Analysis: Disentangled Controls for StyleGAN Image Generationno summary yetcs-cv2011.12799Adobe42 citesNov 25, 2020
- NovCR-Fill: Generative Image Inpainting with Auxiliary Contexutal Reconstructionno summary yetcs-cv2011.12836Adobe0 citesNov 25, 2020
- NovNerfies: Deformable Neural Radiance Fieldsno summary yetcs-cv2011.12948Google Research23 citesNov 25, 2020
- NovGroup-Skeleton-Based Human Action Recognition in Complex Eventsno summary yetcs-cv2011.13273Tencent7 citesNov 26, 2020
- NovSpatio-Temporal Inception Graph Convolutional Networks for Skeleton-Based Action Recognitionno summary yetcs-cv2011.13322Alibaba0 citesNov 26, 2020
- NovGenerative Layout Modeling using Constraint Graphsno summary yetcs-cv2011.13417Adobe9 citesNov 26, 2020
- NovUnsupervised part representation by Flow Capsulesno summary yetcs-cv2011.13920Google Research18 citesNov 27, 2020
- NovRethinking Text Segmentation: A Novel Dataset and A Text-Specific Refinement Approachno summary yetcs-cv2011.14021Adobe2 citesNov 27, 2020
- NovClass-agnostic Object Detectionno summary yetcs-cv2011.14204Amazon0 citesNov 28, 2020
- NovHeuristic Domain Adaptationno summary yetcs-cv2011.14540Alibaba9 citesNov 30, 2020
- NovNeuralFusion: Online Depth Fusion in Latent Spaceno summary yetcs-cv2011.14791Microsoft Research3 citesNov 30, 2020
- NovOne-Shot Free-View Neural Talking-Head Synthesis for Video Conferencingno summary yetcs-cv2011.15126NVIDIA22 citesNov 30, 2020
- NovGenerating Natural Questions from Images for Multimodal Assistantsno summary yetcs-cv2012.03678Apple1 citesNov 17, 2020
- NovUNOC: Understanding Occlusion for Embodied Presence in Virtual Realityno summary yetcs-cv2012.03680Meta / FAIR1 citesNov 12, 2020
- NovPredicting Prostate Cancer-Specific Mortality with A.I.-based Gleason Gradingno summary yetcs-cv2012.05197Google Research49 citesNov 25, 2020
- NovSS-SFDA : Self-Supervised Source-Free Domain Adaptation for Road Segmentation in Hazardous Environmentsno summary yetcs-cv2012.08939Adobe4 citesNov 27, 2020
- OctLearned Dual-View Reflection Removalno summary yetcs-cv2010.00702Adobe0 citesOct 1, 2020
- OctSemi-Supervised Learning for Multi-Task Scene Understanding by Neural Graph Consensusno summary yetcs-cv2010.01086Google Research3 citesOct 2, 2020
- OctAIM 2020 Challenge on Image Extreme Inpaintingno summary yetcs-cv2010.01110Baidu8 citesOct 2, 2020
- OctSemantic MapNet: Building Allocentric Semantic Maps and Representations from Egocentric Viewsno summary yetcs-cv2010.01191Meta / FAIR13 citesOct 2, 2020
- OctConsensus Clustering With Unsupervised Representation Learningno summary yetcs-cv2010.01245Microsoft Research4 citesOct 3, 2020
- OctMagGAN: High-Resolution Face Attribute Editing with Mask-Guided Generative Adversarial Networkno summary yetcs-cv2010.01424Microsoft Research6 citesOct 3, 2020
- OctProbabilistic 3D surface reconstruction from sparse MRI informationno summary yetcs-cv2010.02041Microsoft Research0 citesOct 5, 2020
- OctA Benchmark and Baseline for Language-Driven Image Editingno summary yetcs-cv2010.02330Adobe5 citesOct 5, 2020
- OctRepresentation learning from videos in-the-wild: An object-centric approachno summary yetcs-cv2010.02808Google Research1 citesOct 6, 2020
- OctGlobal Self-Attention Networks for Image Recognitionno summary yetcs-cv2010.03019Google Research20 citesOct 6, 2020
- OctAddressing the Real-world Class Imbalance Problem in Dermatologyno summary yetcs-cv2010.04308Google Research2 citesOct 9, 2020
- OctDeep-Masking Generative Network: A Unified Framework for Background Restoration from Superimposed Imagesno summary yetcs-cv2010.04324Tencent41 citesOct 9, 2020
- OctAccelerate CNNs from Three Dimensions: A Comprehensive Pruning Frameworkno summary yetcs-cv2010.04879Tencent21 citesOct 10, 2020
- OctMulti-path Neural Networks for On-device Multi-domain Visual Classificationno summary yetcs-cv2010.04904Google Research1 citesOct 10, 2020
- OctShape-Texture Debiased Neural Network Trainingno summary yetcs-cv2010.05981Salesforce5 citesOct 12, 2020
- OctMedICaT: A Dataset of Medical Images, Captions, and Textual Referencesno summary yetcs-cv2010.06000AllenAI3 citesOct 12, 2020
- OctKartta Labs: Collaborative Time Travelno summary yetcs-cv2010.06536Google Research0 citesOct 7, 2020
- OctHS-ResNet: Hierarchical-Split Block on Convolutional Neural Networkno summary yetcs-cv2010.07621Baidu35 citesOct 15, 2020
- OctDoes Data Augmentation Benefit from Split BatchNormsno summary yetcs-cv2010.07810Google Research4 citesOct 15, 2020
- OctBoosting Image-based Mutual Gaze Detection using Pseudo 3D Gazeno summary yetcs-cv2010.07811Google Research1 citesOct 15, 2020
- OctVid-ODE: Continuous-Time Video Generation with Neural Ordinary Differential Equationno summary yetcs-cv2010.08188Google Research5 citesOct 16, 2020
- OctDiscovering Pattern Structure Using Differentiable Compositingno summary yetcs-cv2010.08788Adobe1 citesOct 17, 2020
- OctFinding Physical Adversarial Examples for Autonomous Driving with Fast and Differentiable Image Compositingno summary yetcs-cv2010.08844Google Research9 citesOct 17, 2020
- OctDistortion-aware Monocular Depth Estimation for Omnidirectional Imagesno summary yetcs-cv2010.08942Alibaba40 citesOct 18, 2020
- OctGraphite: GRAPH-Induced feaTure Extraction for Point Cloud Registrationno summary yetcs-cv2010.09079Google Research0 citesOct 18, 2020
- OctUnsupervised Domain Adaptation for Spatio-Temporal Action Localizationno summary yetcs-cv2010.09211Google Research4 citesOct 19, 2020
- OctPseudoSeg: Designing Pseudo Labels for Semantic Segmentationno summary yetcs-cv2010.09713Google Research167 citesOct 19, 2020
- OctLT-GAN: Self-Supervised GAN with Latent Transformation Detectionno summary yetcs-cv2010.09893Adobe3 citesOct 19, 2020
- OctSOrT-ing VQA Models : Contrastive Gradient Learning for Improved Consistencyno summary yetcs-cv2010.10038Salesforce0 citesOct 20, 2020
- OctReal-time Localized Photorealistic Video Style Transferno summary yetcs-cv2010.10056Google Research2 citesOct 20, 2020
- OctBiST: Bi-directional Spatio-Temporal Reasoning for Video-Grounded Dialoguesno summary yetcs-cv2010.10095Salesforce1 citesOct 20, 2020
- OctProgressive Batching for Efficient Non-linear Least Squaresno summary yetcs-cv2010.10968Snap0 citesOct 21, 2020
- OctEfficient Scale-Permuted Backbone with Learned Resource Distributionno summary yetcs-cv2010.11426Google Research2 citesOct 22, 2020
- OctFew-Shot Adaptation of Generative Adversarial Networksno summary yetcs-cv2010.11943Google Research54 citesOct 22, 2020
- OctDelving into the Cyclic Mechanism in Semi-supervised Video Object Segmentationno summary yetcs-cv2010.12176Adobe18 citesOct 23, 2020
- OctHard Example Generation by Texture Synthesis for Cross-domain Shape Similarity Learningno summary yetcs-cv2010.12238Alibaba2 citesOct 23, 2020
- OctLoopReg: Self-supervised Learning of Implicit Surface Correspondences, Pose and Shape for 3D Human Mesh Registrationno summary yetcs-cv2010.12447Google Research17 citesOct 23, 2020
- OctView-Invariant, Occlusion-Robust Probabilistic Embedding for Human Poseno summary yetcs-cv2010.13321Google Research4 citesOct 23, 2020
- OctPSF-LO: Parameterized Semantic Features Based Lidar Odometryno summary yetcs-cv2010.13355Alibaba6 citesOct 26, 2020
- OctGreedyFool: Distortion-Aware Sparse Adversarial Attackno summary yetcs-cv2010.13773Microsoft Research12 citesOct 26, 2020
- OctDisplacement-Invariant Matching Cost Learning for Accurate Optical Flow Estimationno summary yetcs-cv2010.14851Tencent9 citesOct 28, 2020
- OctWhy Do Better Loss Functions Lead to Less Transferable Features?no summary yetcs-cv2010.16402Google Research1 citesOct 30, 2020
- OctMichiGAN: Multi-Input-Conditioned Hair Image Generation for Portrait Editingno summary yetcs-cv2010.16417Snap12 citesOct 30, 2020
- SepPIDNet: An Efficient Network for Dynamic Pedestrian Intrusion Detectionno summary yetcs-cv2009.00312Alibaba13 citesSep 1, 2020
- SepText and Style Conditioned GAN for Generation of Offline Handwriting Linesno summary yetcs-cv2009.00678Adobe19 citesSep 1, 2020
- SepSPAN: Spatial Pyramid Attention Network forImage Manipulation Localizationno summary yetcs-cv2009.00726Meta / FAIR235 citesSep 1, 2020
- SepPCPL: Predicate-Correlation Perception Learning for Unbiased Scene Graph Generationno summary yetcs-cv2009.00893Alibaba110 citesSep 2, 2020
- SepMulti-Loss Weighting with Coefficient of Variationsno summary yetcs-cv2009.01717Google Research3 citesSep 3, 2020
- SepFlow-edge Guided Video Completionno summary yetcs-cv2009.01835Meta / FAIR1 citesSep 3, 2020
- SepWitches' Brew: Industrial Scale Data Poisoning via Gradient Matchingno summary yetcs-cv2009.02276Google Research36 citesSep 4, 2020
- SepOne-shot Text Field Labeling using Attention and Belief Propagation for Structure Information Extractionno summary yetcs-cv2009.04153Alibaba0 citesSep 9, 2020
- SepBinarized Neural Architecture Search for Efficient Object Recognitionno summary yetcs-cv2009.04247Baidu4 citesSep 8, 2020
- SepUnderstanding the Role of Individual Units in a Deep Neural Networkno summary yetcs-cv2009.05041Adobe376 citesSep 10, 2020
- SepAttribute-conditioned Layout GAN for Automatic Graphic Designno summary yetcs-cv2009.05284Adobe1 citesSep 11, 2020
- SepDeep Hiearchical Multi-Label Classification Applied to Chest X-Ray Abnormality Taxonomiesno summary yetcs-cv2009.05609NVIDIA0 citesSep 11, 2020
- SepGINet: Graph Interaction Network for Scene Parsingno summary yetcs-cv2009.06160Baidu1 citesSep 14, 2020
- SepBeyond Weak Perspective for Monocular 3D Human Pose Estimationno summary yetcs-cv2009.06549Amazon3 citesSep 14, 2020
- SepAdaptive Text Recognition through Visual Matchingno summary yetcs-cv2009.06610DeepMind1 citesSep 14, 2020
- SepUnderstanding Deformable Alignment in Video Super-Resolutionno summary yetcs-cv2009.07265Tencent17 citesSep 15, 2020
- SepDual Semantic Fusion Network for Video Object Detectionno summary yetcs-cv2009.07498Tencent29 citesSep 16, 2020
- SepAdversarial Image Composition with Auxiliary Illuminationno summary yetcs-cv2009.08255Alibaba8 citesSep 17, 2020
- SepNovel View Synthesis from Single Images via Point Cloud Transformationno summary yetcs-cv2009.08321Google Research1 citesSep 17, 2020
- SepCVPR 2020 Continual Learning in Computer Vision Competition: Approaches, Results, Current Challenges and Future Directionsno summary yetcs-cv2009.09929Google Research10 citesSep 14, 2020
- SepPP-OCR: A Practical Ultra Lightweight OCR Systemno summary yetcs-cv2009.09941Baidu109 citesSep 21, 2020
- SepImproving Person Re-identification with Iterative Impression Aggregationno summary yetcs-cv2009.10066Microsoft Research11 citesSep 21, 2020
- SepPennSyn2Real: Training Object Recognition Models without Human Labelingno summary yetcs-cv2009.10292Amazon0 citesSep 22, 2020
- SepSelf-Supervised Learning of Non-Rigid Residual Flow and Ego-Motionno summary yetcs-cv2009.10467Microsoft Research12 citesSep 22, 2020
- SepDifferential Viewpoints for Ground Terrain Material Recognitionno summary yetcs-cv2009.11072Amazon1 citesSep 22, 2020
- SepX-LXMERT: Paint, Caption and Answer Questions with Multi-Modal Transformersno summary yetcs-cv2009.11278AllenAI24 citesSep 23, 2020
- SepCausal Intervention for Weakly-Supervised Semantic Segmentationno summary yetcs-cv2009.12547Alibaba70 citesSep 26, 2020
- SepLong-Tailed Classification by Keeping the Good and Removing the Bad Momentum Causal Effectno summary yetcs-cv2009.12991Alibaba50 citesSep 28, 2020
- SepRotated Binary Neural Networkno summary yetcs-cv2009.13055Tencent68 citesSep 28, 2020
- SepAsymmetric Loss For Multi-Label Classificationno summary yetcs-cv2009.14119Alibaba64 citesSep 29, 2020
- SepFinding It at Another Side: A Viewpoint-Adapted Matching Encoder for Change Captioningno summary yetcs-cv2009.14352Adobe6 citesSep 30, 2020
- SepPruning Filter in Filterno summary yetcs-cv2009.14410Tencent3 citesSep 30, 2020
- SepMaterialGAN: Reflectance Capture using a Generative SVBRDF Modelno summary yetcs-cv2010.00114Adobe98 citesSep 30, 2020
- SepBoMuDANet: Unsupervised Adaptation for Visual Scene Understanding in Unstructured Driving Environmentsno summary yetcs-cv2010.03523Adobe1 citesSep 22, 2020
- AugSelf-supervised Learning of Point Clouds via Orientation Estimationno summary yetcs-cv2008.00305Adobe7 citesAug 1, 2020
- AugRobust Collaborative Learning of Patch-level and Image-level Annotations for Diabetic Retinopathy Grading from Fundus Imageno summary yetcs-cv2008.00610Baidu4 citesAug 3, 2020
- AugAdversarial Semantic Data Augmentation for Human Pose Estimationno summary yetcs-cv2008.00697Tencent3 citesAug 3, 2020
- AugAE TextSpotter: Learning Visual and Linguistic Representation for Ambiguous Text Spottingno summary yetcs-cv2008.00714Alibaba1 citesAug 3, 2020
- AugWeakly-Supervised Semantic Segmentation via Sub-category Explorationno summary yetcs-cv2008.01183Google Research11 citesAug 3, 2020
- AugPhraseCut: Language-based Image Segmentation in the Wildno summary yetcs-cv2008.01187Adobe1 citesAug 3, 2020
- AugAppearance Consensus Driven Self-Supervised Human Mesh Recoveryno summary yetcs-cv2008.01341Google Research2 citesAug 4, 2020
- AugOpen-Edit: Open-Domain Image Manipulation with Open-Vocabulary Instructionsno summary yetcs-cv2008.01576Adobe7 citesAug 4, 2020
- AugPatchNets: Patch-Based Generalizable Deep Implicit 3D Shape Representationsno summary yetcs-cv2008.01639Meta / FAIR2 citesAug 4, 2020
- AugDeep Multi Depth Panoramas for View Synthesisno summary yetcs-cv2008.01815Adobe1 citesAug 4, 2020
- AugCOALESCE: Component Assembly by Learning to Synthesize Connectionsno summary yetcs-cv2008.01936Adobe6 citesAug 5, 2020
- AugCan You Read Me Now? Content Aware Rectification using Angle Supervisionno summary yetcs-cv2008.02231Amazon2 citesAug 5, 2020
- AugNeRF in the Wild: Neural Radiance Fields for Unconstrained Photo Collectionsno summary yetcs-cv2008.02268Google Research134 citesAug 5, 2020
- AugLearning Illumination from Diverse Portraitsno summary yetcs-cv2008.02396Google Research0 citesAug 5, 2020
- AugStyleFlow: Attribute-conditioned Exploration of StyleGAN-Generated Images using Conditional Continuous Normalizing Flowsno summary yetcs-cv2008.02401Adobe388 citesAug 6, 2020
- AugLearning to Factorize and Relight a Cityno summary yetcs-cv2008.02796Google Research3 citesAug 6, 2020
- AugPredicting Visual Importance Across Graphic Design Typesno summary yetcs-cv2008.02912Adobe5 citesAug 7, 2020
- AugMulti-Level Temporal Pyramid Network for Action Detectionno summary yetcs-cv2008.03270Alibaba2 citesAug 7, 2020
- AugFeature Space Augmentation for Long-Tailed Datano summary yetcs-cv2008.03673Google Research13 citesAug 9, 2020
- AugNeural Light Transport for Relighting and View Synthesisno summary yetcs-cv2008.03806Google Research77 citesAug 9, 2020
- AugNeural Reflectance Fields for Appearance Acquisitionno summary yetcs-cv2008.03824Adobe110 citesAug 9, 2020
- AugText as Neural Operator: Image Manipulation by Text Instructionno summary yetcs-cv2008.04556Google Research28 citesAug 11, 2020
- AugSharp Multiple Instance Learning for DeepFake Video Detectionno summary yetcs-cv2008.04585Alibaba152 citesAug 11, 2020
- AugReal-Time Sign Language Detection using Human Pose Estimationno summary yetcs-cv2008.04637Google Research10 citesAug 11, 2020
- AugGeLaTO: Generative Latent Textured Objectsno summary yetcs-cv2008.04852Google Research3 citesAug 11, 2020
- AugAdversarial Generative Grammars for Human Activity Predictionno summary yetcs-cv2008.04888Google Research1 citesAug 11, 2020
- AugImage segmentation via Cellular Automatano summary yetcs-cv2008.04965Google Research8 citesAug 11, 2020
- AugLearning to Caricature via Semantic Shape Transformno summary yetcs-cv2008.05090Tencent3 citesAug 12, 2020
- AugLook here! A parametric learning based approach to redirect visual attentionno summary yetcs-cv2008.05413Adobe0 citesAug 12, 2020
- AugWhat leads to generalization of object proposals?no summary yetcs-cv2008.05700Meta / FAIR2 citesAug 13, 2020
- AugLift, Splat, Shoot: Encoding Images From Arbitrary Camera Rigs by Implicitly Unprojecting to 3Dno summary yetcs-cv2008.05711NVIDIA41 citesAug 13, 2020
- AugAdversarial Knowledge Transfer from Unlabeled Datano summary yetcs-cv2008.05746Adobe0 citesAug 13, 2020
- AugGeoLayout: Geometry Driven Room Layout Estimation Based on Depth Maps of Planesno summary yetcs-cv2008.06286Google Research3 citesAug 14, 2020
- AugAntiDote: Attention-based Dynamic Optimization for Neural Network Runtime Efficiencyno summary yetcs-cv2008.06543Microsoft Research0 citesAug 14, 2020
- AugPoet: Product-oriented Video Captioner for E-commerceno summary yetcs-cv2008.06880Alibaba25 citesAug 16, 2020
- AugDeVLBert: Learning Deconfounded Visio-Linguistic Representationsno summary yetcs-cv2008.06884Alibaba63 citesAug 16, 2020
- AugNeural Descent for Visual 3D Human Pose and Shapeno summary yetcs-cv2008.06910Google Research6 citesAug 16, 2020
- AugDo Not Disturb Me: Person Re-identification Under the Interference of Other Pedestriansno summary yetcs-cv2008.06963Tencent8 citesAug 16, 2020
- AugAP-Loss for Accurate One-Stage Object Detectionno summary yetcs-cv2008.07294Tencent87 citesAug 17, 2020
- AugSoftPoolNet: Shape Descriptor for Point Cloud Completion and Classificationno summary yetcs-cv2008.07358Google Research5 citesAug 17, 2020
- AugZero Shot Domain Generalizationno summary yetcs-cv2008.07443Microsoft Research1 citesAug 17, 2020
- AugPix2Surf: Learning Parametric 3D Surface Models of Objects from Imagesno summary yetcs-cv2008.07760Adobe1 citesAug 18, 2020
- AugAssembleNet++: Assembling Modality Representations via Attention Connectionsno summary yetcs-cv2008.08072Google Research8 citesAug 18, 2020
- AugPC-U Net: Learning to Jointly Reconstruct and Segment the Cardiac Walls in 3D from CT Datano summary yetcs-cv2008.08194NVIDIA22 citesAug 18, 2020
- AugDeepHandMesh: A Weakly-supervised Deep Encoder-Decoder Framework for High-fidelity Hand Mesh Modelingno summary yetcs-cv2008.08213Meta / FAIR4 citesAug 19, 2020
- AugCFAD: Coarse-to-Fine Action Detector for Spatiotemporal Action Localizationno summary yetcs-cv2008.08332Adobe0 citesAug 19, 2020
- AugAutoSimulate: (Quickly) Learning Synthetic Data Generationno summary yetcs-cv2008.08424Microsoft Research2 citesAug 16, 2020
- AugDeepGMR: Learning Latent Gaussian Mixture Models for Registrationno summary yetcs-cv2008.09088NVIDIA16 citesAug 20, 2020
- AugInterHand2.6M: A Dataset and Baseline for 3D Interacting Hand Pose Estimation from a Single RGB Imageno summary yetcs-cv2008.09309Meta / FAIR13 citesAug 21, 2020
- AugToward Quantifying Ambiguities in Artistic Imagesno summary yetcs-cv2008.09688Adobe0 citesAug 21, 2020
- AugThe Hessian Penalty: A Weak Prior for Unsupervised Disentanglementno summary yetcs-cv2008.10599Adobe11 citesAug 24, 2020
- AugAdaptive Context-Aware Multi-Modal Network for Depth Completionno summary yetcs-cv2008.10833Alibaba167 citesAug 25, 2020
- AugWeakly Supervised Learning with Side Information for Noisy Labeled Imagesno summary yetcs-cv2008.11586Alibaba5 citesAug 25, 2020
- AugLearning Global Structure Consistency for Robust Object Trackingno summary yetcs-cv2008.11769Baidu2 citesAug 26, 2020
- AugBackground Splitting: Finding Rare Classes in a Sea of Backgroundno summary yetcs-cv2008.12873Google Research0 citesAug 28, 2020
- AugFinding Action Tubes with a Sparse-to-Dense Frameworkno summary yetcs-cv2008.13196Adobe0 citesAug 30, 2020
- AugSelf-supervised Video Representation Learning by Uncovering Spatio-temporal Statisticsno summary yetcs-cv2008.13426Tencent23 citesAug 31, 2020
- JulAttention-Oriented Action Recognition for Real-Time Human-Robot Interactionno summary yetcs-cv2007.01065Tencent2 citesJul 2, 2020
- JulSegment as Points for Efficient Online Multi-Object Tracking and Segmentationno summary yetcs-cv2007.01550Baidu2 citesJul 3, 2020
- JulLearning to Discover Multi-Class Attentional Regions for Multi-Label Image Recognitionno summary yetcs-cv2007.01755Tencent130 citesJul 3, 2020
- JulImproving Weakly Supervised Visual Grounding by Contrastive Knowledge Distillationno summary yetcs-cv2007.01951Tencent8 citesJul 3, 2020
- JulDessiLBI: Exploring Structural Sparsity of Deep Networks via Differential Inclusion Pathsno summary yetcs-cv2007.02010Microsoft Research0 citesJul 4, 2020
- JulRobust Processing-In-Memory Neural Networks via Noise-Aware Normalizationno summary yetcs-cv2007.03230Google Research10 citesJul 7, 2020
- JulReal-time Semantic Segmentation with Fast Attentionno summary yetcs-cv2007.03815Meta / FAIR6 citesJul 7, 2020
- JulRemix: Rebalanced Mixupno summary yetcs-cv2007.03943Google Research57 citesJul 8, 2020
- JulAligning Videos in Space and Timeno summary yetcs-cv2007.04515Meta / FAIR1 citesJul 9, 2020
- JulGeneralized Few-Shot Video Classification with Video Retrieval and Feature Generationno summary yetcs-cv2007.04755Meta / FAIR0 citesJul 9, 2020
- JulPIE-NET: Parametric Inference of Point Cloud Edgesno summary yetcs-cv2007.04883Google Research43 citesJul 9, 2020
- JulScientific Discovery by Generating Counterfactuals using Image Translationno summary yetcs-cv2007.05500Google Research2 citesJul 10, 2020
- JulMulti-Domain Image Completion for Random Missing Input Datano summary yetcs-cv2007.05534NVIDIA5 citesJul 10, 2020
- JulLearning and Exploiting Interclass Visual Correlations for Medical Image Classificationno summary yetcs-cv2007.06371Tencent1 citesJul 13, 2020
- JulTowards causal benchmarking of bias in face analysis algorithmsno summary yetcs-cv2007.06570Amazon6 citesJul 13, 2020
- JulAUTO3D: Novel view synthesis through unsupervisely learned variational viewpoint and global 3D representationno summary yetcs-cv2007.06620Meta / FAIR25 citesJul 13, 2020
- JulJNR: Joint-based Neural Rig Representation for Compact 3D Face Modelingno summary yetcs-cv2007.06755Microsoft Research0 citesJul 14, 2020
- JulCollaborative Unsupervised Domain Adaptation for Medical Image Diagnosisno summary yetcs-cv2007.07222Tencent186 citesJul 5, 2020
- JulModeling Artistic Workflows for Image Generation and Editingno summary yetcs-cv2007.07238Adobe1 citesJul 14, 2020
- JulA Generalization of Otsu's Method and Minimum Error Thresholdingno summary yetcs-cv2007.07350Google Research6 citesJul 14, 2020
- JulComparing to Learn: Surpassing ImageNet Pretraining on Radiographs By Comparing Image Representationsno summary yetcs-cv2007.07423Tencent11 citesJul 15, 2020
- JulAttention-Based Query Expansion Learningno summary yetcs-cv2007.08019Meta / FAIR4 citesJul 15, 2020
- JulLearning End-to-End Action Interaction by Paired-Embedding Data Augmentationno summary yetcs-cv2007.08071Tencent1 citesJul 16, 2020
- JulControllable Image Synthesis via SegVAEno summary yetcs-cv2007.08397Google Research2 citesJul 16, 2020
- JulRetrieveGAN: Image Synthesis via Differentiable Patch Retrievalno summary yetcs-cv2007.08513Google Research4 citesJul 16, 2020
- JulInfoFocus: 3D Object Detection for Autonomous Driving with Dynamic Information Modelingno summary yetcs-cv2007.08556Salesforce2 citesJul 16, 2020
- JulOn Robustness and Transferability of Convolutional Neural Networksno summary yetcs-cv2007.08558Google Research23 citesJul 16, 2020
- JulAdvances in Deep Learning for Hyperspectral Image Analysis--Addressing Challenges Arising in Practical Imaging Scenariosno summary yetcs-cv2007.08592Amazon10 citesJul 16, 2020
- JulAdaptive Task Sampling for Meta-Learningno summary yetcs-cv2007.08735Salesforce5 citesJul 17, 2020
- JulDVI: Depth Guided Video Inpainting for Autonomous Drivingno summary yetcs-cv2007.08854Baidu2 citesJul 17, 2020
- JulConsensus-Aware Visual-Semantic Embedding for Image-Text Matchingno summary yetcs-cv2007.08883Tencent17 citesJul 17, 2020
- JulGeometric Correspondence Fields: Learned Differentiable Rendering for 3D Pose Refinement in the Wildno summary yetcs-cv2007.08939Meta / FAIR1 citesJul 17, 2020
- JulGenerating Person Images with Appearance-aware Pose Stylizerno summary yetcs-cv2007.09077Baidu3 citesJul 17, 2020
- JulImproving Object Detection with Selective Self-supervised Self-trainingno summary yetcs-cv2007.09162Google Research1 citesJul 17, 2020
- JulSpeech2Video Synthesis with 3D Skeleton Regularization and Expressive Body Posesno summary yetcs-cv2007.09198Baidu5 citesJul 17, 2020
- JulFeature Pyramid Transformerno summary yetcs-cv2007.09451Alibaba7 citesJul 18, 2020
- JulSingle View Metrology in the Wildno summary yetcs-cv2007.09529Adobe1 citesJul 18, 2020
- JulContactPose: A Dataset of Grasps with Object Contact and Hand Poseno summary yetcs-cv2007.09545Meta / FAIR14 citesJul 19, 2020
- JulSeeing the Un-Scene: Learning Amodal Semantic Maps for Room Navigationno summary yetcs-cv2007.09841Meta / FAIR6 citesJul 20, 2020
- JulInterpretable Foreground Object Search As Knowledge Distillationno summary yetcs-cv2007.09867Alibaba1 citesJul 20, 2020
- JulDeep Reflectance Volumes: Relightable Reconstructions from Multi-View Photometric Imagesno summary yetcs-cv2007.09892Adobe17 citesJul 20, 2020
- JulIncorporating Reinforced Adversarial Learning in Autoregressive Image Generationno summary yetcs-cv2007.09923Adobe1 citesJul 20, 2020
- JulGREEN: a Graph REsidual rE-ranking Network for Grading Diabetic Retinopathyno summary yetcs-cv2007.09968Tencent1 citesJul 20, 2020
- JulDistractor-Aware Neuron Intrinsic Learning for Generic 2D Medical Image Classificationsno summary yetcs-cv2007.09979Tencent0 citesJul 20, 2020
- JulDeep Image Clustering with Category-Style Representationno summary yetcs-cv2007.10004Tencent8 citesJul 20, 2020
- JulA Macro-Micro Weakly-supervised Framework for AS-OCT Tissue Segmentationno summary yetcs-cv2007.10007Tencent1 citesJul 20, 2020
- JulCoupling Explicit and Implicit Surface Representations for Generative 3D Modelingno summary yetcs-cv2007.10294Adobe5 citesJul 20, 2020
- JulJoint Disentangling and Adaptation for Cross-Domain Person Re-Identificationno summary yetcs-cv2007.10315NVIDIA19 citesJul 20, 2020
- JulPillar-based Object Detection for Autonomous Drivingno summary yetcs-cv2007.10323Google Research29 citesJul 20, 2020
- JulUnified Multisensory Perception: Weakly-Supervised Audio-Visual Video Parsingno summary yetcs-cv2007.10558Adobe9 citesJul 21, 2020
- JulGraph-PCNN: Two Stage Human Pose Estimation with Graph Pose Refinementno summary yetcs-cv2007.10599Baidu12 citesJul 21, 2020
- JulMulti-modal Transformer for Video Retrievalno summary yetcs-cv2007.10639Google Research0 citesJul 21, 2020
- JulUncertainty-Aware Weakly Supervised Action Detection from Untrimmed Videosno summary yetcs-cv2007.10703Google Research5 citesJul 21, 2020
- JulDense Hybrid Recurrent Multi-view Stereo Net with Dynamic Consistency Checkingno summary yetcs-cv2007.10872Tencent10 citesJul 21, 2020
- JulPointContrast: Unsupervised Pre-training for 3D Point Cloud Understandingno summary yetcs-cv2007.10985Meta / FAIR95 citesJul 21, 2020
- JulInstance-aware Self-supervised Learning for Nuclei Segmentationno summary yetcs-cv2007.11186Tencent4 citesJul 22, 2020
- JulLearnable Cost Volume Using the Cayley Representationno summary yetcs-cv2007.11431Google Research1 citesJul 21, 2020
- JulCombining Implicit Function Learning and Parametric Models for 3D Human Reconstructionno summary yetcs-cv2007.11432Google Research13 citesJul 22, 2020
- JulCrossTransformers: spatially-aware few-shot transferno summary yetcs-cv2007.11498Google Research58 citesJul 22, 2020
- JulSIZER: A Dataset and Model for Parsing 3D Clothing and Learning Size Sensitive 3D Clothingno summary yetcs-cv2007.11610Meta / FAIR5 citesJul 22, 2020
- JulContact and Human Dynamics from Monocular Videono summary yetcs-cv2007.11678Adobe0 citesJul 22, 2020
- JulEnd-to-end Learning of Compressible Featuresno summary yetcs-cv2007.11797Google Research3 citesJul 23, 2020
- JulNeural Geometric Parser for Single Image Camera Calibrationno summary yetcs-cv2007.11855Adobe0 citesJul 23, 2020
- JulSBAT: Video Captioning with Sparse Boundary-Aware Transformerno summary yetcs-cv2007.11888Baidu3 citesJul 23, 2020
- JulThe Devil is in Classification: A Simple Framework for Long-tail Object Detection and Instance Segmentationno summary yetcs-cv2007.11978Salesforce4 citesJul 23, 2020
- JulUnsupervised Deep Representation Learning for Real-Time Trackingno summary yetcs-cv2007.11984Tencent7 citesJul 22, 2020
- JulAttentionNAS: Spatiotemporal Attention Cell Search for Video Classificationno summary yetcs-cv2007.12034Google Research7 citesJul 23, 2020
- JulPP-YOLO: An Effective and Efficient Implementation of Object Detectorno summary yetcs-cv2007.12099Baidu234 citesJul 23, 2020
- JulHITNet: Hierarchical Iterative Tile Refinement Network for Real-time Stereo Matchingno summary yetcs-cv2007.12140Google Research10 citesJul 23, 2020
- JulSpatially Aware Multimodal Transformers for TextVQAno summary yetcs-cv2007.12146Meta / FAIR2 citesJul 23, 2020
- JulAn LSTM Approach to Temporal 3D Object Detection in LiDAR Point Cloudsno summary yetcs-cv2007.12392Google Research11 citesJul 24, 2020
- JulFully Convolutional Networks for Continuous Sign Language Recognitionno summary yetcs-cv2007.12402Tencent145 citesJul 24, 2020
- JulLearning Crisp Edge Detector Using Logical Refinement Networkno summary yetcs-cv2007.12449Tencent1 citesJul 24, 2020
- JulRobust and Generalizable Visual Representation Learning via Random Convolutionsno summary yetcs-cv2007.13003Google Research16 citesJul 25, 2020
- JulMask2CAD: 3D Shape Prediction by Learning to Segment and Retrieveno summary yetcs-cv2007.13034Google Research6 citesJul 26, 2020
- JulVirtual Multi-view Fusion for 3D Semantic Segmentationno summary yetcs-cv2007.13138Google Research13 citesJul 26, 2020
- JulNOH-NMS: Improving Pedestrian Detection by Nearby Objects Hallucinationno summary yetcs-cv2007.13376Tencent4 citesJul 27, 2020
- JulDeep Hashing with Hash-Consistent Large Margin Proxy Embeddingsno summary yetcs-cv2007.13912Netflix0 citesJul 27, 2020
- JulActive Learning for Video Description With Cluster-Regularized Ensemble Rankingno summary yetcs-cv2007.13913Google Research1 citesJul 27, 2020
- JulAccurate, Low-Latency Visual Perception for Autonomous Racing:Challenges, Mechanisms, and Practical Solutionsno summary yetcs-cv2007.13971DeepMind0 citesJul 28, 2020
- JulLearning Modality Interaction for Temporal Sentence Localization and Event Captioning in Videosno summary yetcs-cv2007.14164Tencent11 citesJul 28, 2020
- JulDive Deeper Into Box for Object Detectionno summary yetcs-cv2007.14350Tencent1 citesJul 15, 2020
- JulChained-Tracker: Chaining Paired Attentive Regression Results for End-to-End Joint Multiple-Object Detection and Trackingno summary yetcs-cv2007.14557Tencent32 citesJul 29, 2020
- JulLearning Video Representations from Textual Web Supervisionno summary yetcs-cv2007.14937Google Research16 citesJul 29, 2020
- JulFully Dynamic Inference with Deep Neural Networksno summary yetcs-cv2007.15151Meta / FAIR1 citesJul 29, 2020
- JulSimPose: Effectively Learning DensePose and Surface Normals of People from Simulated Datano summary yetcs-cv2007.15506Google Research3 citesJul 30, 2020
- JulRewriting a Deep Generative Modelno summary yetcs-cv2007.15646Adobe4 citesJul 30, 2020
- JulPerceiving 3D Human-Object Spatial Arrangements from a Single Image in the Wildno summary yetcs-cv2007.15649Meta / FAIR6 citesJul 30, 2020
- JulContrastive Learning for Unpaired Image-to-Image Translationno summary yetcs-cv2007.15651Adobe126 citesJul 30, 2020
- JulWeakly supervised one-stage vision and language disease detection using large scale pneumonia and pneumothorax studiesno summary yetcs-cv2007.15778NVIDIA1 citesJul 31, 2020
- JulKAPLAN: A 3D Point Descriptor for Shape Completionno summary yetcs-cv2008.00096Microsoft Research2 citesJul 31, 2020
- JunChannel Attention based Iterative Residual Learning for Depth Map Super-Resolutionno summary yetcs-cv2006.01469Baidu7 citesJun 2, 2020
- JunBlack-box Explanation of Object Detectors via Saliency Mapsno summary yetcs-cv2006.03204Adobe11 citesJun 5, 2020
- JunFP-Stereo: Hardware-Efficient Stereo Vision for Embedded Applicationsno summary yetcs-cv2006.03250Alibaba2 citesJun 5, 2020
- JunAutoHAS: Efficient Hyperparameter and Architecture Searchno summary yetcs-cv2006.03656Google Research25 citesJun 5, 2020
- JunWOAD: Weakly Supervised Online Action Detection in Untrimmed Videosno summary yetcs-cv2006.03732Salesforce4 citesJun 5, 2020
- JunAssociate-3Ddet: Perceptual-to-Conceptual Association for 3D Point Cloud Object Detectionno summary yetcs-cv2006.04356Baidu15 citesJun 8, 2020
- JunGeneralized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object Detectionno summary yetcs-cv2006.04388Microsoft Research272 citesJun 8, 2020
- JunWhat Matters in Unsupervised Optical Flowno summary yetcs-cv2006.04902Google Research8 citesJun 8, 2020
- JunDialog Policy Learning for Joint Clarification and Active Learning Queriesno summary yetcs-cv2006.05456Amazon1 citesJun 9, 2020
- JunRevisiting visual-inertial structure from motion for odometry and SLAM initializationno summary yetcs-cv2006.06017Snap0 citesJun 10, 2020
- JunTelling Left from Right: Learning Spatial Correspondence of Sight and Soundno summary yetcs-cv2006.06175Adobe2 citesJun 11, 2020
- JunRethinking Pre-training and Self-trainingno summary yetcs-cv2006.06882Google Research49 citesJun 11, 2020
- JunWeakly-supervised Temporal Action Localization by Uncertainty Modelingno summary yetcs-cv2006.07006Microsoft Research10 citesJun 12, 2020
- JunCascaded deep monocular 3D human pose estimation with evolutionary training datano summary yetcs-cv2006.07778Tencent169 citesJun 14, 2020
- JunShapeFlow: Learnable Deformations Among 3D Shapesno summary yetcs-cv2006.07982Google Research14 citesJun 14, 2020
- JunGeo-PIFu: Geometry and Pixel Aligned Implicit Functions for Single-view Human Reconstructionno summary yetcs-cv2006.08072Adobe29 citesJun 15, 2020
- JunMulti-Image Summarization: Textual Summary from a Set of Cohesive Imagesno summary yetcs-cv2006.08686Google Research1 citesJun 15, 2020
- JunMetaSDF: Meta-learning Signed Distance Functionsno summary yetcs-cv2006.09662Google Research29 citesJun 17, 2020
- JunUnsupervised Learning of Visual Features by Contrasting Cluster Assignmentsno summary yetcs-cv2006.09882Meta / FAIR1,909 citesJun 17, 2020
- JunBlazePose: On-device Real-time Body Pose trackingno summary yetcs-cv2006.10204Google Research86 citesJun 17, 2020
- JunMediaPipe Hands: On-device Real-time Hand Trackingno summary yetcs-cv2006.10214Google Research552 citesJun 18, 2020
- JunTowards a Neural Graphics Pipeline for Controllable Image Generationno summary yetcs-cv2006.10569Adobe1 citesJun 18, 2020
- JunDiverse Image Generation via Self-Conditioned GANsno summary yetcs-cv2006.10728Adobe2 citesJun 18, 2020
- JunSpin-Weighted Spherical CNNsno summary yetcs-cv2006.10731Google Research10 citesJun 18, 2020
- JunDifferentiable Augmentation for Data-Efficient GAN Trainingno summary yetcs-cv2006.10738Adobe100 citesJun 18, 2020
- JunAttention Mesh: High-fidelity Face Mesh Prediction in Real-timeno summary yetcs-cv2006.10962Google Research51 citesJun 19, 2020
- JunConsistency Guided Scene Flow Estimationno summary yetcs-cv2006.11242Google Research0 citesJun 19, 2020
- JunVideo Panoptic Segmentationno summary yetcs-cv2006.11339Adobe1 citesJun 19, 2020
- JunReal-time Pupil Tracking from Monocular Video for Digital Puppetryno summary yetcs-cv2006.11341Google Research18 citesJun 19, 2020
- JunLAMP: Large Deep Nets with Automated Model Parallelism for Image Segmentationno summary yetcs-cv2006.12575NVIDIA6 citesJun 22, 2020
- JunFNA++: Fast Network Adaptation via Parameter Remapping and Architecture Searchno summary yetcs-cv2006.12986Google Research4 citesJun 21, 2020
- JunInstant 3D Object Tracking with Applications in Augmented Realityno summary yetcs-cv2006.13194Google Research9 citesJun 23, 2020
- JunDISK: Learning local features with policy gradientno summary yetcs-cv2006.13566Google Research184 citesJun 24, 2020
- JunRetrospective Loss: Looking Back to Improve Training of Deep Neural Networksno summary yetcs-cv2006.13593Adobe0 citesJun 24, 2020
- JunComprehensive Information Integration Modeling Framework for Video Titlingno summary yetcs-cv2006.13608Alibaba23 citesJun 24, 2020
- JunLayoutTransformer: Layout Generation and Completion with Self-attentionno summary yetcs-cv2006.14615Amazon6 citesJun 25, 2020
- JunPerspective Plane Program Induction from a Single Imageno summary yetcs-cv2006.14708Google Research1 citesJun 25, 2020
- JunUnsupervised Video Decomposition using Spatio-temporal Iterative Inferenceno summary yetcs-cv2006.14727Google Research7 citesJun 25, 2020
- JunCross-Supervised Object Detectionno summary yetcs-cv2006.15056Google Research4 citesJun 26, 2020
- JunCounting Out Time: Class Agnostic Video Repetition Counting in the Wildno summary yetcs-cv2006.15418DeepMind3 citesJun 27, 2020
- JunJoint Hand-object 3D Reconstruction from a Single Image with Cross-branch Feature Fusionno summary yetcs-cv2006.15561Tencent73 citesJun 28, 2020
- JunSelf-Supervised MultiModal Versatile Networksno summary yetcs-cv2006.16228Google Research195 citesJun 29, 2020
- JunUncertainty-aware multi-view co-training for semi-supervised medical image segmentation and domain adaptationno summary yetcs-cv2006.16806NVIDIA263 citesJun 28, 2020
- JunConditional Set Generation with Transformersno summary yetcs-cv2006.16841Google Research24 citesJun 26, 2020
- MayThe AVA-Kinetics Localized Human Actions Video Datasetno summary yetcs-cv2005.00214Google Research84 citesMay 1, 2020
- MayTransforming and Projecting Images into Class-conditional Generative Networksno summary yetcs-cv2005.01703Adobe6 citesMay 4, 2020
- MayStreaming Object Detection for 3-D Point Cloudsno summary yetcs-cv2005.01864Google Research6 citesMay 4, 2020
- MayFrom Image Collections to Point Clouds with Self-supervised Shape and Pose Networksno summary yetcs-cv2005.01939Google Research2 citesMay 5, 2020
- MayUnsupervised Instance Segmentation in Microscopy Images via Panoptic Domain Adaptation and Task Re-weightingno summary yetcs-cv2005.02066Microsoft Research4 citesMay 5, 2020
- MayAGE Challenge: Angle Closure Glaucoma Evaluation in Anterior Segment Optical Coherence Tomographyno summary yetcs-cv2005.02258Baidu9 citesMay 5, 2020
- MayCascadePSP: Toward Class-Agnostic and Very High-Resolution Segmentation via Global and Local Refinementno summary yetcs-cv2005.02551Tencent17 citesMay 6, 2020
- MaySurfelGAN: Synthesizing Realistic Sensor Data for Autonomous Drivingno summary yetcs-cv2005.03844Google Research4 citesMay 8, 2020
- MayVectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representationno summary yetcs-cv2005.04259Google Research44 citesMay 8, 2020
- MayEpipolar Transformersno summary yetcs-cv2005.04551Meta / FAIR0 citesMay 10, 2020
- MayA Simple Semi-Supervised Learning Framework for Object Detectionno summary yetcs-cv2005.04757Google Research326 citesMay 10, 2020
- MayA Closed-Form Uncertainty Propagation in Non-Rigid Structure from Motionno summary yetcs-cv2005.04810NVIDIA6 citesMay 10, 2020
- MayPrototypical Contrastive Learning of Unsupervised Representationsno summary yetcs-cv2005.04966Salesforce484 citesMay 11, 2020
- MayHiFaceGAN: Face Renovation via Collaborative Suppression and Replenishmentno summary yetcs-cv2005.05005Alibaba132 citesMay 11, 2020
- MayEffective and Robust Detection of Adversarial Examples via Benford-Fourier Coefficientsno summary yetcs-cv2005.05552Tencent8 citesMay 12, 2020
- MayAmbient Sound Helps: Audiovisual Crowd Counting in Extreme Conditionsno summary yetcs-cv2005.07097Baidu30 citesMay 14, 2020
- MayTaskology: Utilizing Task Relations at Scaleno summary yetcs-cv2005.07289Google Research3 citesMay 14, 2020
- MayHistory for Visual Dialog: Do we really need it?no summary yetcs-cv2005.07493Adobe8 citesMay 8, 2020
- MayFace Identity Disentanglement via Latent Space Mappingno summary yetcs-cv2005.07728Alibaba10 citesMay 15, 2020
- MayT-VSE: Transformer-Based Visual Semantic Embeddingno summary yetcs-cv2005.08399Google Research4 citesMay 17, 2020
- MayCharacter Matters: Video Story Understanding with Character-Aware Relationsno summary yetcs-cv2005.08646Amazon12 citesMay 9, 2020
- MayMMFashion: An Open-Source Toolbox for Visual Fashion Analysisno summary yetcs-cv2005.08847Alibaba11 citesMay 18, 2020
- MayPortrait Shadow Manipulationno summary yetcs-cv2005.08925Google Research2 citesMay 18, 2020
- MayOn the Value of Out-of-Distribution Testing: An Example of Goodhart's Lawno summary yetcs-cv2005.09241Adobe18 citesMay 19, 2020
- MayDifferentiable Mapping Networks: Learning Structured Map Representations for Sparse Visual Localizationno summary yetcs-cv2005.09530Google Research2 citesMay 19, 2020
- MayWeakly Supervised Representation Learning with Coarse Labelsno summary yetcs-cv2005.09681Alibaba0 citesMay 19, 2020
- MayActive Speakers in Contextno summary yetcs-cv2005.09812Adobe0 citesMay 20, 2020
- MayRange Conditioned Dilated Convolutions for Scale Invariant 3D Object Detectionno summary yetcs-cv2005.09927Google Research12 citesMay 20, 2020
- MayDynamic Refinement Network for Oriented and Densely Packed Object Detectionno summary yetcs-cv2005.09973Tencent14 citesMay 20, 2020
- MayWhat Makes for Good Views for Contrastive Learning?no summary yetcs-cv2005.10243Google Research183 citesMay 20, 2020
- MayNaive-Student: Leveraging Semi-Supervised Learning in Video Sequences for Urban Scene Segmentationno summary yetcs-cv2005.10266Google Research12 citesMay 20, 2020
- MayAOWS: Adaptive and optimal network width search with latency constraintsno summary yetcs-cv2005.10481Amazon3 citesMay 21, 2020
- MayBenefits of temporal information for appearance-based gaze estimationno summary yetcs-cv2005.11670Meta / FAIR15 citesMay 24, 2020
- MayHigh-Resolution Image Inpainting with Iterative Confidence Feedback and Guided Upsamplingno summary yetcs-cv2005.11742Adobe7 citesMay 24, 2020
- MayArbitrary Style Transfer via Multi-Adaptation Networkno summary yetcs-cv2005.13219Tencent10 citesMay 27, 2020
- MayPredicting Goal-directed Human Attention Using Inverse Reinforcement Learningno summary yetcs-cv2005.14310Adobe1 citesMay 28, 2020
- MayUGC-VQA: Benchmarking Blind Video Quality Assessment for User Generated Contentno summary yetcs-cv2005.14354Google Research293 citesMay 29, 2020
- AprCurricularFace: Adaptive Curriculum Learning Loss for Deep Face Recognitionno summary yetcs-cv2004.00288Tencent30 citesApr 1, 2020
- AprPIFuHD: Multi-Level Pixel-Aligned Implicit Function for High-Resolution 3D Human Digitizationno summary yetcs-cv2004.00452Meta / FAIR42 citesApr 1, 2020
- AprEvading Deepfake-Image Detectors with White- and Black-Box Attacksno summary yetcs-cv2004.00622Google Research14 citesApr 1, 2020
- AprImproving 3D Object Detection through Progressive Population Based Augmentationno summary yetcs-cv2004.00831Google Research12 citesApr 2, 2020
- AprMCEN: Bridging Cross-Modal Gap between Cooking Recipes and Dish Images with Latent Variable Modelno summary yetcs-cv2004.01095Alibaba2 citesApr 2, 2020
- AprDOPS: Learning to Detect 3D Objects and Predict their 3D Shapesno summary yetcs-cv2004.01170Google Research8 citesApr 2, 2020
- AprLearning to See Through Obstructionsno summary yetcs-cv2004.01180Google Research2 citesApr 2, 2020
- AprDeformation-Aware 3D Model Embedding and Retrievalno summary yetcs-cv2004.01228Adobe4 citesApr 2, 2020
- AprNovel View Synthesis of Dynamic Scenes with Globally Coherent Depths from a Monocular Camerano summary yetcs-cv2004.01294NVIDIA0 citesApr 2, 2020
- AprLiDAR-based Online 3D Video Object Detection with Graph-based Message Passing and Spatiotemporal Transformer Attentionno summary yetcs-cv2004.01389Baidu14 citesApr 3, 2020
- AprGradient Centralization: A New Optimization Technique for Deep Neural Networksno summary yetcs-cv2004.01461Alibaba32 citesApr 3, 2020
- AprSqueezeSegV3: Spatially-Adaptive Convolution for Efficient Point-Cloud Segmentationno summary yetcs-cv2004.01803Meta / FAIR28 citesApr 3, 2020
- AprGoogle Landmarks Dataset v2 -- A Large-Scale Benchmark for Instance-Level Recognition and Retrievalno summary yetcs-cv2004.01804Google Research28 citesApr 3, 2020
- AprDeblurring by Realistic Blurringno summary yetcs-cv2004.01860Tencent11 citesApr 4, 2020
- AprSimAug: Learning Robust Representations from Simulation for Trajectory Predictionno summary yetcs-cv2004.02022Google Research15 citesApr 4, 2020
- AprDeep Homography Estimation for Dynamic Scenesno summary yetcs-cv2004.02132Adobe2 citesApr 5, 2020
- AprBiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentationno summary yetcs-cv2004.02147Tencent115 citesApr 5, 2020
- AprLightweight Multi-View 3D Pose Estimation through Camera-Disentangled Representationno summary yetcs-cv2004.02186Meta / FAIR9 citesApr 5, 2020
- AprSteering Self-Supervised Feature Learning Beyond Local Pixel Statisticsno summary yetcs-cv2004.02331Adobe5 citesApr 5, 2020
- AprAutoToon: Automatic Geometric Warping for Face Cartoon Generationno summary yetcs-cv2004.02377Adobe0 citesApr 6, 2020
- AprDeep Space-Time Video Upsampling Networksno summary yetcs-cv2004.02432Meta / FAIR0 citesApr 6, 2020
- AprGANSpace: Discovering Interpretable GAN Controlsno summary yetcs-cv2004.02546Adobe425 citesApr 6, 2020
- AprBeyond the Nav-Graph: Vision-and-Language Navigation in Continuous Environmentsno summary yetcs-cv2004.02857Meta / FAIR26 citesApr 6, 2020
- AprObjectness-Aware Few-Shot Semantic Segmentationno summary yetcs-cv2004.02945Adobe5 citesApr 6, 2020
- AprLearning Generative Models of Shape Handlesno summary yetcs-cv2004.03028Adobe2 citesApr 6, 2020
- AprHuman Motion Transfer from Poses in the Wildno summary yetcs-cv2004.03142Snap10 citesApr 7, 2020
- AprMotion-supervised Co-Part Segmentationno summary yetcs-cv2004.03234Snap8 citesApr 7, 2020
- AprAttribution in Scale and Spaceno summary yetcs-cv2004.03383Google Research3 citesApr 3, 2020
- AprPatchVAE: Learning Local Latent Codes for Recognitionno summary yetcs-cv2004.03623Google Research2 citesApr 7, 2020
- AprSemantic Image Manipulation Using Scene Graphsno summary yetcs-cv2004.03677Google Research2 citesApr 7, 2020
- AprContext-Aware Group Captioning via Self-Attention and Contrastive Featuresno summary yetcs-cv2004.03708Adobe3 citesApr 7, 2020
- AprState of the Art on Neural Renderingno summary yetcs-cv2004.03805Adobe19 citesApr 8, 2020
- AprLearning 3D Semantic Scene Graphs from 3D Indoor Reconstructionsno summary yetcs-cv2004.03967Google Research10 citesApr 8, 2020
- AprThe GeoLifeCLEF 2020 Datasetno summary yetcs-cv2004.04192Microsoft Research2 citesApr 8, 2020
- AprSelf-Supervised 3D Human Pose Estimation via Part Guided Novel Image Synthesisno summary yetcs-cv2004.04400Google Research5 citesApr 9, 2020
- AprScalable Active Learning for Object Detectionno summary yetcs-cv2004.04699NVIDIA4 citesApr 9, 2020
- AprImproving Semantic Segmentation through Spatio-Temporal Consistency Learned from Videosno summary yetcs-cv2004.05324Google Research1 citesApr 11, 2020
- AprCompositional Visual Generation and Inference with Energy Based Modelsno summary yetcs-cv2004.06030Google Research8 citesApr 13, 2020
- AprSpeedNet: Learning the Speediness in Videosno summary yetcs-cv2004.06130Google Research21 citesApr 13, 2020
- AprEmbedded Large-Scale Handwritten Chinese Character Recognitionno summary yetcs-cv2004.06209Apple6 citesApr 13, 2020
- AprIntuitive, Interactive Beard and Hair Synthesis with Generative Modelsno summary yetcs-cv2004.06848Tencent2 citesApr 15, 2020
- AprDeep-COVID: Predicting COVID-19 From Chest X-Ray Images Using Deep Transfer Learningno summary yetcs-cv2004.09363Snap60 citesApr 20, 2020
- AprPanoptic-based Image Synthesisno summary yetcs-cv2004.10289NVIDIA3 citesApr 21, 2020
- AprEfficient Neighbourhood Consensus Networks via Submanifold Sparse Convolutionsno summary yetcs-cv2004.10566DeepMind5 citesApr 22, 2020
- AprVisualCOMET: Reasoning about the Dynamic Context of a Still Imageno summary yetcs-cv2004.10796AllenAI12 citesApr 22, 2020
- AprJoint Bilateral Learning for Real-time Universal Photorealistic Style Transferno summary yetcs-cv2004.10955Google Research5 citesApr 23, 2020
- AprSingle-View View Synthesis with Multiplane Imagesno summary yetcs-cv2004.11364Google Research5 citesApr 23, 2020
- AprMining self-similarity: Label super-resolution with epitomic representationsno summary yetcs-cv2004.11498Microsoft Research0 citesApr 24, 2020
- AprAny Motion Detector: Learning Class-agnostic Scene Dynamics from a Sequence of LiDAR Point Cloudsno summary yetcs-cv2004.11647Google Research5 citesApr 24, 2020
- AprLearning to Autofocusno summary yetcs-cv2004.12260Google Research0 citesApr 26, 2020
- AprFashionpedia: Ontology, Segmentation, and an Attribute Localization Datasetno summary yetcs-cv2004.12276Google Research4 citesApr 26, 2020
- AprCoReNet: Coherent 3D scene reconstruction from a single RGB imageno summary yetcs-cv2004.12989Google Research4 citesApr 27, 2020
- AprMakeItTalk: Speaker-Aware Talking-Head Animationno summary yetcs-cv2004.12992Adobe308 citesApr 27, 2020
- AprVD-BERT: A Unified Vision and Dialog Transformer with BERTno summary yetcs-cv2004.13278Salesforce30 citesApr 28, 2020
- AprNeural Hair Renderingno summary yetcs-cv2004.13297Snap0 citesApr 28, 2020
- AprMulti-Scale Boosted Dehazing Network with Dense Feature Fusionno summary yetcs-cv2004.13388Google Research63 citesApr 28, 2020
- AprLeveraging Photometric Consistency over Time for Sparsely Supervised Hand-Object Reconstructionno summary yetcs-cv2004.13449Microsoft Research3 citesApr 28, 2020
- AprMotion Guided 3D Pose Estimation from Videosno summary yetcs-cv2004.13985Amazon10 citesApr 29, 2020
- AprDetecting Deep-Fake Videos from Appearance and Behaviorno summary yetcs-cv2004.14491Meta / FAIR25 citesApr 29, 2020
- AprThe 4th AI City Challengeno summary yetcs-cv2004.14619Amazon0 citesApr 30, 2020
- AprSimPropNet: Improved Similarity Propagation for Few-shot Image Segmentationno summary yetcs-cv2004.15014Adobe8 citesApr 30, 2020
- MarCops-Ref: A new Dataset and Task on Compositional Referring Expression Comprehensionno summary yetcs-cv2003.00403Tencent6 citesMar 1, 2020
- MarPointASNL: Robust Point Clouds Processing using Nonlocal Neural Networks with Adaptive Samplingno summary yetcs-cv2003.00492Tencent49 citesMar 1, 2020
- MarZoomNet: Part-Aware Adaptive Zooming Neural Network for 3D Object Detectionno summary yetcs-cv2003.00529Baidu4 citesMar 1, 2020
- MarTowards Noise-resistant Object Detection with Noisy Annotationsno summary yetcs-cv2003.01285Salesforce15 citesMar 3, 2020
- MarUnsupervised Learning of Intrinsic Structural Representation Pointsno summary yetcs-cv2003.01661Adobe3 citesMar 3, 2020
- MarRegion adaptive graph fourier transform for 3d point cloudsno summary yetcs-cv2003.01866Google Research0 citesMar 4, 2020
- MarGarmentGAN: Photo-realistic Adversarial Fashion Transferno summary yetcs-cv2003.01894Salesforce12 citesMar 4, 2020
- MarCreating High Resolution Images with a Latent Adversarial Generatorno summary yetcs-cv2003.02365Google Research10 citesMar 4, 2020
- MarDA4AD: End-to-End Deep Attention-based Visual Localization for Autonomous Drivingno summary yetcs-cv2003.03026Baidu2 citesMar 6, 2020
- MarSemi-Supervised StyleGAN for Disentanglement Learningno summary yetcs-cv2003.03461Amazon14 citesMar 6, 2020
- MarPoseNet3D: Learning Temporally Consistent 3D Human Pose via Knowledge Distillationno summary yetcs-cv2003.03473Amazon2 citesMar 7, 2020
- MarMobilePose: Real-Time Pose Estimation for Unseen Objects with Weak Shape Supervisionno summary yetcs-cv2003.03522Google Research27 citesMar 7, 2020
- MarContext-Aware Domain Adaptation in Semantic Segmentationno summary yetcs-cv2003.04010Tencent4 citesMar 9, 2020
- MarCascaded Human-Object Interaction Recognitionno summary yetcs-cv2003.04262Google Research8 citesMar 9, 2020
- MarVisual Grounding in Video for Unsupervised Word Translationno summary yetcs-cv2003.05078DeepMind6 citesMar 11, 2020
- MarDual Temporal Memory Network for Efficient Video Object Segmentationno summary yetcs-cv2003.06125Netflix1 citesMar 13, 2020
- MarSemantic Pyramid for Image Generationno summary yetcs-cv2003.06221Google Research3 citesMar 13, 2020
- MarExplainable Deep Classification Models for Domain Generalizationno summary yetcs-cv2003.06498Adobe5 citesMar 13, 2020
- MarNoiseRank: Unsupervised Label Noise Reduction with Dependence Modelsno summary yetcs-cv2003.06729Meta / FAIR7 citesMar 15, 2020
- MarClosed-loop Matters: Dual Regression Networks for Single Image Super-Resolutionno summary yetcs-cv2003.07018Baidu32 citesMar 16, 2020
- MarLT-Net: Label Transfer by Learning Reversible Voxel-wise Correspondence for One-shot Medical Image Segmentationno summary yetcs-cv2003.07072Tencent5 citesMar 16, 2020
- MarNeural Pose Transfer by Spatially Adaptive Instance Normalizationno summary yetcs-cv2003.07254Google Research7 citesMar 16, 2020
- MarParameter-Free Style Projection for Arbitrary Style Transferno summary yetcs-cv2003.07694Baidu5 citesMar 17, 2020
- MarAxial-DeepLab: Stand-Alone Axial-Attention for Panoptic Segmentationno summary yetcs-cv2003.07853Google Research66 citesMar 17, 2020
- MarWatching the World Go By: Representation Learning from Unlabeled Videosno summary yetcs-cv2003.07990AllenAI38 citesMar 18, 2020
- MarLighthouse: Predicting Lighting Volumes for Spatially-Coherent Illuminationno summary yetcs-cv2003.08367Google Research13 citesMar 18, 2020
- MarPairwise Similarity Knowledge Transfer for Weakly Supervised Object Localizationno summary yetcs-cv2003.08375Google Research2 citesMar 18, 2020
- MarCollaborative Distillation for Ultra-Resolution Universal Style Transferno summary yetcs-cv2003.08436Adobe7 citesMar 18, 2020
- MarA Metric Learning Reality Checkno summary yetcs-cv2003.08505Meta / FAIR67 citesMar 18, 2020
- MarSAPIEN: A SimulAted Part-based Interactive ENvironmentno summary yetcs-cv2003.08515Google Research22 citesMar 19, 2020
- MarPose Augmentation: Class-agnostic Object Pose Transformation for Object Recognitionno summary yetcs-cv2003.08526Google Research0 citesMar 19, 2020
- MarLocal Implicit Grid Representations for 3D Scenesno summary yetcs-cv2003.08981Google Research65 citesMar 19, 2020
- MarAffinity Graph Supervision for Visual Recognitionno summary yetcs-cv2003.09049Adobe0 citesMar 19, 2020
- MarData-Free Knowledge Amalgamation via Group-Stack Dual-GANno summary yetcs-cv2003.09088Alibaba6 citesMar 20, 2020
- MarWeakly Supervised 3D Hand Pose Estimation via Biomechanical Constraintsno summary yetcs-cv2003.09282NVIDIA14 citesMar 20, 2020
- MarLearning 3D Part Assembly from a Single Imageno summary yetcs-cv2003.09754Adobe2 citesMar 21, 2020
- MarLifespan Age Transformation Synthesisno summary yetcs-cv2003.09764Adobe8 citesMar 21, 2020
- MarThe Instantaneous Accuracy: a Novel Metric for the Problem of Online Human Behaviour Recognition in Untrimmed Videosno summary yetcs-cv2003.09970Adobe2 citesMar 22, 2020
- MarLearning Better Lossless Compression Using Lossy Compressionno summary yetcs-cv2003.10184Google Research3 citesMar 23, 2020
- MarWeakly Supervised 3D Human Pose and Shape Reconstruction with Normalizing Flowsno summary yetcs-cv2003.10350Google Research11 citesMar 23, 2020
- MarGen-LaneNet: A Generalized and Scalable Approach for 3D Lane Detectionno summary yetcs-cv2003.10656Baidu6 citesMar 24, 2020
- MarRethinking Class-Balanced Methods for Long-Tailed Visual Recognition from a Domain Adaptation Perspectiveno summary yetcs-cv2003.10780Google Research29 citesMar 24, 2020
- MarDeep Local Shapes: Learning Local SDF Priors for Detailed 3D Reconstructionno summary yetcs-cv2003.10983Meta / FAIR17 citesMar 24, 2020
- MarBigNAS: Scaling Up Neural Architecture Search with Big Single-Stage Modelsno summary yetcs-cv2003.11142Google Research38 citesMar 24, 2020
- MarA Unified Object Motion and Affinity Model for Online Multi-Object Trackingno summary yetcs-cv2003.11291Baidu3 citesMar 25, 2020
- MarImproved Techniques for Training Single-Image GANsno summary yetcs-cv2003.11512Adobe9 citesMar 25, 2020
- MarRethinking Few-Shot Image Classification: a Good Embedding Is All You Need?no summary yetcs-cv2003.11539Google Research88 citesMar 25, 2020
- MarDeepStrip: High Resolution Boundary Refinementno summary yetcs-cv2003.11670Adobe1 citesMar 25, 2020
- MarTowards Backward-Compatible Representation Learningno summary yetcs-cv2003.11942Amazon5 citesMar 26, 2020
- MarMilking CowMask for Semi-Supervised Image Classificationno summary yetcs-cv2003.12022Google Research19 citesMar 26, 2020
- MarAre Labels Necessary for Neural Architecture Search?no summary yetcs-cv2003.12056Meta / FAIR5 citesMar 26, 2020
- MarGrounded Situation Recognitionno summary yetcs-cv2003.12058AllenAI0 citesMar 26, 2020
- MarParSeNet: A Parametric Surface Fitting Network for 3D Point Cloudsno summary yetcs-cv2003.12181Adobe9 citesMar 26, 2020
- MarTextCaps: a Dataset for Image Captioning with Reading Comprehensionno summary yetcs-cv2003.12462Meta / FAIR19 citesMar 24, 2020
- MarDA-NAS: Data Adapted Pruning for Efficient Neural Architecture Searchno summary yetcs-cv2003.12563Microsoft Research2 citesMar 27, 2020
- MarDeep 3D Capture: Geometry and Reflectance from Sparse Multi-View Imagesno summary yetcs-cv2003.12642Adobe7 citesMar 27, 2020
- MarDeep CG2Real: Synthetic-to-Real Translation via Image Disentanglementno summary yetcs-cv2003.12649Adobe0 citesMar 27, 2020
- MarNMS by Representative Region: Towards Crowded Pedestrian Detection by Proposal Pairingno summary yetcs-cv2003.12729Tencent31 citesMar 28, 2020
- MarSuperpixel Segmentation with Fully Convolutional Networksno summary yetcs-cv2003.12929Adobe16 citesMar 29, 2020
- MarCross-Domain Document Object Detection: Benchmark Suite and Methodno summary yetcs-cv2003.13197Adobe5 citesMar 30, 2020
- MarArchitecture Disentanglement for Deep Neural Networksno summary yetcs-cv2003.13268Tencent2 citesMar 30, 2020
- MarContext Based Emotion Recognition using EMOTIC Datasetno summary yetcs-cv2003.13401NVIDIA206 citesMar 30, 2020
- MarSpeech2Action: Cross-modal Supervision for Action Recognitionno summary yetcs-cv2003.13594DeepMind11 citesMar 30, 2020
- MarUnderstanding the impact of mistakes on background regions in crowd countingno summary yetcs-cv2003.13759Amazon2 citesMar 30, 2020
- Mar3D-MPA: Multi Proposal Aggregation for 3D Semantic Instance Segmentationno summary yetcs-cv2003.13867Google Research31 citesMar 30, 2020
- MarRetinaTrack: Online Single Stage Joint Detection and Trackingno summary yetcs-cv2003.13870Google Research15 citesMar 30, 2020
- MarNeural Networks Are More Productive Teachers Than Human Raters: Active Mixup for Data-Efficient Knowledge Distillation from a Blackbox Modelno summary yetcs-cv2003.13960Google Research4 citesMar 31, 2020
- MarFaceScape: a Large-scale High Quality 3D Face Dataset and Detailed Riggable 3D Face Predictionno summary yetcs-cv2003.13989Baidu12 citesMar 31, 2020
- MarReal-Time Semantic Segmentation via Auto Depth, Downsampling Joint Decision and Feature Aggregationno summary yetcs-cv2003.14226Tencent1 citesMar 31, 2020
- MarDu$^2$Net: Learning Depth Estimation from Dual-Cameras and Dual-Pixelsno summary yetcs-cv2003.14299Google Research2 citesMar 31, 2020
- MarDeep Semantic Matching with Foreground Detection and Cycle-Consistencyno summary yetcs-cv2004.00144Google Research0 citesMar 31, 2020
- MarExploring Long Tail Visual Relationship Recognition with Large Vocabularyno summary yetcs-cv2004.00436Baidu0 citesMar 25, 2020
- FebAdversarially Robust Frame Sampling with Bounded Irregularitiesno summary yetcs-cv2002.01147Google Research0 citesFeb 4, 2020
- FebRevisiting Spatial Invariance with Low-Rank Local Connectivityno summary yetcs-cv2002.02959Google Research10 citesFeb 7, 2020
- FebDynamic Inference: A New Approach Toward Efficient Video Action Recognitionno summary yetcs-cv2002.03342Baidu3 citesFeb 9, 2020
- FebImproving Face Recognition from Hard Samples via Distribution Distillation Lossno summary yetcs-cv2002.03662Tencent3 citesFeb 10, 2020
- FebWhy Do Line Drawings Work? A Realism Hypothesisno summary yetcs-cv2002.06260Adobe56 citesFeb 14, 2020
- FebVideo Face Super-Resolution with Motion-Adaptive Feedback Cellno summary yetcs-cv2002.06378Tencent1 citesFeb 15, 2020
- FebFacial Attribute Capsules for Noise Face Super Resolutionno summary yetcs-cv2002.06518Tencent1 citesFeb 16, 2020
- FebPrecision Gating: Improving Neural Network Efficiency with Dynamic Dual-Precision Activationsno summary yetcs-cv2002.07136Microsoft Research2 citesFeb 17, 2020
- FebDivideMix: Learning with Noisy Labels as Semi-supervised Learningno summary yetcs-cv2002.07394Salesforce89 citesFeb 18, 2020
- FebWeakly-Supervised Semantic Segmentation by Iterative Affinity Learningno summary yetcs-cv2002.08098Tencent87 citesFeb 19, 2020
- FebA Neural Lip-Sync Framework for Synthesizing Photorealistic Virtual News Anchorsno summary yetcs-cv2002.08700Google Research2 citesFeb 20, 2020
- FebAutomatic Shortcut Removal for Self-Supervised Representation Learningno summary yetcs-cv2002.08822Google Research36 citesFeb 20, 2020
- FebBlockGAN: Learning 3D Object-aware Scene Representations from Unlabelled Imagesno summary yetcs-cv2002.08988Adobe93 citesFeb 20, 2020
- FebSketchformer: Transformer-based Representation for Sketched Structureno summary yetcs-cv2002.10381Adobe30 citesFeb 24, 2020
- FebTowards Learning a Generic Agent for Vision-and-Language Navigation via Pre-trainingno summary yetcs-cv2002.10638Microsoft Research18 citesFeb 25, 2020
- FebSuper-Resolving Commercial Satellite Imagery Using Realistic Training Datano summary yetcs-cv2002.11248Google Research1 citesFeb 26, 2020
- FebAdversarial Ranking Attack and Defenseno summary yetcs-cv2002.11293Alibaba0 citesFeb 26, 2020
- FebEfficient Semantic Video Segmentation with Per-frame Inferenceno summary yetcs-cv2002.11433Microsoft Research9 citesFeb 26, 2020
- FebEvolving Losses for Unsupervised Video Representation Learningno summary yetcs-cv2002.12177Google Research6 citesFeb 26, 2020
- FebCross-modality Person re-identification with Shared-Specific Feature Transferno summary yetcs-cv2002.12489Alibaba23 citesFeb 28, 2020
- FebNAS-Count: Counting-by-Density with Neural Architecture Searchno summary yetcs-cv2003.00217Alibaba5 citesFeb 29, 2020
- FebMulti-Scale Representation Learning for Spatial Feature Distributions using Grid Cellsno summary yetcs-cv2003.00824LinkedIn15 citesFeb 16, 2020
- JanA Multi-oriented Chinese Keyword Spotter Guided by Text Line Detectionno summary yetcs-cv2001.00722Tencent0 citesJan 3, 2020
- JanDiscrimination-aware Network Pruning for Deep Model Compressionno summary yetcs-cv2001.01050Tencent144 citesJan 4, 2020
- JanAn Exploration of Embodied Visual Explorationno summary yetcs-cv2001.02192Meta / FAIR7 citesJan 7, 2020
- JanFast Neural Network Adaptation via Parameter Remapping and Architecture Searchno summary yetcs-cv2001.02525Google Research15 citesJan 8, 2020
- JanAn Analysis of Object Representations in Deep Visual Trackersno summary yetcs-cv2001.02593Google Research0 citesJan 8, 2020
- JanDon't Judge an Object by Its Context: Learning to Overcome Contextual Biasno summary yetcs-cv2001.03152Meta / FAIR8 citesJan 9, 2020
- JanPruning Convolutional Neural Networks with Self-Supervisionno summary yetcs-cv2001.03554Meta / FAIR31 citesJan 10, 2020
- JanRetouchdown: Adding Touchdown to StreetLearn as a Shareable Resource for Language Grounding Tasks in Street Viewno summary yetcs-cv2001.03671Google Research20 citesJan 10, 2020
- JanFine-grained Image-to-Image Transformation towards Visual Recognitionno summary yetcs-cv2001.03856Tencent1 citesJan 12, 2020
- JanHigh-Fidelity Synthesis with Disentangled Representationno summary yetcs-cv2001.04296Google Research16 citesJan 13, 2020
- JanRoutedFusion: Learning Real-time Depth Map Fusionno summary yetcs-cv2001.04388Microsoft Research5 citesJan 13, 2020
- JanEGO-TOPO: Environment Affordances from Egocentric Videono summary yetcs-cv2001.04583Meta / FAIR5 citesJan 14, 2020
- JanNeural Architecture Search for Deep Image Priorno summary yetcs-cv2001.04776Adobe8 citesJan 14, 2020
- JanUnifying Deep Local and Global Features for Image Searchno summary yetcs-cv2001.05027Google Research34 citesJan 14, 2020
- JanProposal Learning for Semi-Supervised Object Detectionno summary yetcs-cv2001.05086Salesforce11 citesJan 15, 2020
- JanImage Segmentation Using Deep Learning: A Surveyno summary yetcs-cv2001.05566Snap143 citesJan 15, 2020
- JanLE-HGR: A Lightweight and Efficient RGB-based Online Gesture Recognition Network for Embedded AR Devicesno summary yetcs-cv2001.05654Alibaba7 citesJan 16, 2020
- JanFilter Grafting for Deep Neural Networksno summary yetcs-cv2001.05868Tencent9 citesJan 15, 2020
- JanSieveNet: A Unified Framework for Robust Image-Based Virtual Try-Onno summary yetcs-cv2001.06265Adobe5 citesJan 17, 2020
- JanMulti-View Photometric Stereo: A Robust Solution and Benchmark Dataset for Spatially Varying Isotropic Materialsno summary yetcs-cv2001.06659Tencent1 citesJan 18, 2020
- JanGated Path Selection Network for Semantic Segmentationno summary yetcs-cv2001.06819Baidu28 citesJan 19, 2020
- JanFD-GAN: Generative Adversarial Networks with Fusion-discriminator for Single Image Dehazingno summary yetcs-cv2001.06968Adobe1 citesJan 20, 2020
- JanLearning Diverse Features with Part-Level Resolution for Person Re-Identificationno summary yetcs-cv2001.07442Alibaba6 citesJan 21, 2020
- JanOptimized Generic Feature Learning for Few-shot Classification across Domainsno summary yetcs-cv2001.07926Google Research29 citesJan 22, 2020
- JanDetecting Deficient Coverage in Colonoscopiesno summary yetcs-cv2001.08589Google Research2 citesJan 23, 2020
- JanSOLAR: Second-Order Loss and Attention for Image Retrievalno summary yetcs-cv2001.08972Meta / FAIR12 citesJan 24, 2020
- JanVerSe: A Vertebrae Labelling and Segmentation Benchmark for Multi-detector CT Imagesno summary yetcs-cv2001.09193Alibaba345 citesJan 24, 2020
- JanCurriculum Audiovisual Learningno summary yetcs-cv2001.09414Baidu33 citesJan 26, 2020
- JanUnsupervised Disentanglement of Pose, Appearance and Background from Images and Videosno summary yetcs-cv2001.09518NVIDIA5 citesJan 26, 2020
- JanFast Video Object Segmentation using the Global Context Moduleno summary yetcs-cv2001.11243Tencent8 citesJan 30, 2020
- JanLearn to Predict Sets Using Feed-Forward Neural Networksno summary yetcs-cv2001.11845Amazon3 citesJan 30, 2020
- JanConvolutional Hierarchical Attention Network for Query-Focused Video Summarizationno summary yetcs-cv2002.03740Alibaba65 citesJan 31, 2020
2019
397- DecModeling Affect-based Intrinsic Rewards for Exploration and Learningno summary yetcs-cv1912.00403Microsoft Research1 citesDec 1, 2019
- DecLatentFusion: End-to-End Differentiable Reconstruction and Rendering for Unseen Object Pose Estimationno summary yetcs-cv1912.00416NVIDIA6 citesDec 1, 2019
- DecInformation bottleneck through variational glassesno summary yetcs-cv1912.00830Google Research23 citesDec 2, 2019
- DecA Multigrid Method for Efficiently Training Video Modelsno summary yetcs-cv1912.00998Meta / FAIR5 citesDec 2, 2019
- DecView-Invariant Probabilistic Embedding for Human Poseno summary yetcs-cv1912.01001Google Research4 citesDec 2, 2019
- DecAsymmetric Co-Teaching for Unsupervised Cross Domain Person Re-Identificationno summary yetcs-cv1912.01349Tencent13 citesDec 3, 2019
- DecIt GAN DO Better: GAN-based Detection of Objects on Images with Varying Qualityno summary yetcs-cv1912.01707Apple9 citesDec 3, 2019
- DecTowards Robust Image Classification Using Sequential Attention Modelsno summary yetcs-cv1912.02184DeepMind9 citesDec 4, 2019
- Dec12-in-1: Multi-Task Vision and Language Representation Learningno summary yetcs-cv1912.02315Meta / FAIR36 citesDec 5, 2019
- DecUltrafast Photorealistic Style Transfer via Neural Architecture Searchno summary yetcs-cv1912.02398Baidu5 citesDec 5, 2019
- DecSelf-Supervised Learning of Video-Induced Visual Invariancesno summary yetcs-cv1912.02783Google Research17 citesDec 5, 2019
- DecKeyPose: Multi-View 3D Labeling and Keypoint Estimation for Transparent Objectsno summary yetcs-cv1912.02805Google Research7 citesDec 5, 2019
- DecWhy Having 10,000 Parameters in Your Camera Model is Better Than Twelveno summary yetcs-cv1912.02908Microsoft Research4 citesDec 5, 2019
- DecPyramid Multi-view Stereo Net with Self-adaptive View Aggregationno summary yetcs-cv1912.03001Tencent8 citesDec 6, 2019
- DecConnecting Vision and Language with Localized Narrativesno summary yetcs-cv1912.03098Google Research25 citesDec 6, 2019
- DecContext R-CNN: Long Term Temporal Context for Per-Camera Object Detectionno summary yetcs-cv1912.03538Google Research6 citesDec 7, 2019
- DecVoronoiNet: General Functional Approximators with Local Supportno summary yetcs-cv1912.03629Google Research1 citesDec 8, 2019
- DecShape-Aware Organ Segmentation by Predicting Signed Distance Mapsno summary yetcs-cv1912.03849Tencent13 citesDec 9, 2019
- DecLearning a Neural 3D Texture Space from 2D Exemplarsno summary yetcs-cv1912.04158Adobe3 citesDec 9, 2019
- DecGrasping in the Wild:Learning 6DoF Closed-Loop Grasping from Low-Cost Demonstrationsno summary yetcs-cv1912.04344Google Research9 citesDec 9, 2019
- DecBasis Prediction Networks for Effective Burst Denoising with Large Kernelsno summary yetcs-cv1912.04421Adobe7 citesDec 9, 2019
- DecNeural Voxel Renderer: Learning an Accurate and Controllable Rendering Toolno summary yetcs-cv1912.04591Google Research2 citesDec 10, 2019
- DecNeural Point Cloud Rendering via Multi-Plane Projectionno summary yetcs-cv1912.04645Google Research8 citesDec 10, 2019
- DecSpineNet: Learning Scale-Permuted Backbone for Recognition and Localizationno summary yetcs-cv1912.05027Google Research21 citesDec 10, 2019
- DeccFineGAN: Unsupervised multi-conditional fine-grained image generationno summary yetcs-cv1912.05028Adobe0 citesDec 6, 2019
- DecLearning from Noisy Anchors for One-stage Object Detectionno summary yetcs-cv1912.05086Salesforce6 citesDec 11, 2019
- DecLocal Context Normalization: Revisiting Local Normalizationno summary yetcs-cv1912.05845Microsoft Research4 citesDec 12, 2019
- DecLocal Deep Implicit Functions for 3D Shapeno summary yetcs-cv1912.06126Google Research5 citesDec 12, 2019
- DecSim2Real Predictivity: Does Evaluation in Simulation Predict Real-World Performance?no summary yetcs-cv1912.06321Meta / FAIR157 citesDec 13, 2019
- DecEnd-to-End Learning of Visual Representations from Uncurated Instructional Videosno summary yetcs-cv1912.06430DeepMind39 citesDec 13, 2019
- DecThe Garden of Forking Paths: Towards Multi-Future Trajectory Predictionno summary yetcs-cv1912.06445Google Research11 citesDec 13, 2019
- DecViBE: Dressing for Diverse Body Shapesno summary yetcs-cv1912.06697Meta / FAIR2 citesDec 13, 2019
- DecCross-Modality Attention with Semantic Graph Embedding for Multi-Label Classificationno summary yetcs-cv1912.07872Baidu16 citesDec 17, 2019
- DecTowards Generalization Across Depth for Monocular 3D Object Detectionno summary yetcs-cv1912.08035Meta / FAIR7 citesDec 17, 2019
- DecLearning Generalizable Visual Representations via Interactive Gameplayno summary yetcs-cv1912.08195AllenAI12 citesDec 17, 2019
- DecCoupled Network for Robust Pedestrian Detection with Gated Multi-Layer Feature Extraction and Deformable Occlusion Handlingno summary yetcs-cv1912.08661Tencent2 citesDec 18, 2019
- DecSynSin: End-to-end View Synthesis from a Single Imageno summary yetcs-cv1912.08804Meta / FAIR15 citesDec 18, 2019
- DecSimulating Content Consistent Vehicle Datasets with Attribute Descentno summary yetcs-cv1912.08855NVIDIA32 citesDec 18, 2019
- DecNeural Design Network: Graphic Layout Generation with Constraintsno summary yetcs-cv1912.09421Google Research9 citesDec 19, 2019
- DecLearning Semantic Neural Tree for Human Parsingno summary yetcs-cv1912.09622Tencent3 citesDec 20, 2019
- DecAtomNAS: Fine-Grained End-to-End Neural Architecture Searchno summary yetcs-cv1912.09640Adobe42 citesDec 20, 2019
- DecDeepSFM: Structure From Motion Via Deep Bundle Adjustmentno summary yetcs-cv1912.09697Google Research4 citesDec 20, 2019
- DecAdvanced Variations of Two-Dimensional Principal Component Analysis for Face Recognitionno summary yetcs-cv1912.09970Baidu3 citesDec 19, 2019
- DecJacobian Adversarially Regularized Networks for Robustnessno summary yetcs-cv1912.10185Google Research8 citesDec 21, 2019
- DecFasterSeg: Searching for Faster Real-time Semantic Segmentationno summary yetcs-cv1912.10917Google Research58 citesDec 23, 2019
- DecCNN-generated images are surprisingly easy to spot... for nowno summary yetcs-cv1912.11035Adobe43 citesDec 23, 2019
- DecBig Transfer (BiT): General Visual Representation Learningno summary yetcs-cv1912.11370Google Research179 citesDec 24, 2019
- DecSoundSpaces: Audio-Visual Navigation in 3D Environmentsno summary yetcs-cv1912.11474Meta / FAIR7 citesDec 24, 2019
- DecLook, Listen, and Act: Towards Audio-Visual Embodied Navigationno summary yetcs-cv1912.11684Google Research10 citesDec 25, 2019
- DecControllable and Progressive Image Extrapolationno summary yetcs-cv1912.11711Adobe9 citesDec 25, 2019
- DecCategory-Level Articulated Object Pose Estimationno summary yetcs-cv1912.11913Google Research7 citesDec 26, 2019
- DecMachine Learning for Precipitation Nowcasting from Radar Imagesno summary yetcs-cv1912.12132Google Research240 citesDec 11, 2019
- DecLocality and compositionality in zero-shot learningno summary yetcs-cv1912.12179Microsoft Research21 citesDec 20, 2019
- DecExplain Your Move: Understanding Agent Actions Using Specific and Relevant Feature Attributionno summary yetcs-cv1912.12191Adobe12 citesDec 23, 2019
- DecHybrid Channel Based Pedestrian Detectionno summary yetcs-cv1912.12431Alibaba33 citesDec 28, 2019
- DecDepthTransfer: Depth Extraction from Video Using Non-parametric Samplingno summary yetcs-cv2001.00987Microsoft Research495 citesDec 24, 2019
- DecDepth Extraction from Video Using Non-parametric Samplingno summary yetcs-cv2002.04479Microsoft Research274 citesDec 24, 2019
- NovDD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Framesno summary yetcs-cv1911.00357Google Research36 citesNov 1, 2019
- NovSelf-supervised Deformation Modeling for Facial Expression Editingno summary yetcs-cv1911.00735Adobe0 citesNov 2, 2019
- NovHigh Fidelity Video Prediction with Large Stochastic Recurrent Neural Networksno summary yetcs-cv1911.01655Adobe20 citesNov 5, 2019
- NovRoIMix: Proposal-Fusion among Multiple Images for Underwater Object Detectionno summary yetcs-cv1911.03029Tencent16 citesNov 8, 2019
- NovLearning Deep Bilinear Transformation for Fine-grained Image Representationno summary yetcs-cv1911.03621Microsoft Research52 citesNov 9, 2019
- NovCSPN++: Learning Context and Resource Aware Convolutional Spatial Propagation Networks for Depth Completionno summary yetcs-cv1911.05377Baidu15 citesNov 13, 2019
- NovSelf-Supervised Learning For Few-Shot Image Classificationno summary yetcs-cv1911.06045Alibaba26 citesNov 14, 2019
- NovSemantic Granularity Metric Learning for Visual Searchno summary yetcs-cv1911.06047Amazon3 citesNov 14, 2019
- NovLearning To Characterize Adversarial Subspacesno summary yetcs-cv1911.06587Alibaba0 citesNov 15, 2019
- NovIn-domain representation learning for remote sensingno summary yetcs-cv1911.06721Google Research34 citesNov 15, 2019
- NovLabel-similarity Curriculum Learningno summary yetcs-cv1911.06902Microsoft Research6 citesNov 15, 2019
- NovDynamic Instance Normalization for Arbitrary Style Transferno summary yetcs-cv1911.06953Baidu18 citesNov 16, 2019
- NovBSP-Net: Generating Compact Meshes via Binary Space Partitioningno summary yetcs-cv1911.06971Google Research13 citesNov 16, 2019
- NovImprove CAM with Auto-adapted Segmentation and Co-supervised Augmentationno summary yetcs-cv1911.07160Adobe1 citesNov 17, 2019
- NovDistribution Context Aware Loss for Person Re-identificationno summary yetcs-cv1911.07273Alibaba1 citesNov 17, 2019
- NovLarge Scale Open-Set Deep Logo Detectionno summary yetcs-cv1911.07440Google Research2 citesNov 18, 2019
- NovDermGAN: Synthetic Generation of Clinical Skin Images with Pathologyno summary yetcs-cv1911.08716Google Research44 citesNov 20, 2019
- NovRefineDetLite: A Lightweight One-stage Object Detection Framework for CPU-only Devicesno summary yetcs-cv1911.08855Tencent28 citesNov 20, 2019
- NovEfficientDet: Scalable and Efficient Object Detectionno summary yetcs-cv1911.09070Google Research526 citesNov 20, 2019
- NovThe Origins and Prevalence of Texture Bias in Convolutional Neural Networksno summary yetcs-cv1911.09071Google Research120 citesNov 20, 2019
- NovSearch to Distill: Pearls are Everywhere but not the Eyesno summary yetcs-cv1911.09074Google Research3 citesNov 20, 2019
- NovMulti-Label Classification with Label Graph Superimposingno summary yetcs-cv1911.09243Baidu14 citesNov 21, 2019
- NovAdversarial Examples Improve Image Recognitionno summary yetcs-cv1911.09665Google Research48 citesNov 21, 2019
- NovFast Sparse ConvNetsno summary yetcs-cv1911.09723DeepMind12 citesNov 21, 2019
- NovReinforcing an Image Caption Generator Using Off-Line Human Feedbackno summary yetcs-cv1911.09753Google Research3 citesNov 21, 2019
- NovPanoptic-DeepLab: A Simple, Strong, and Fast Baseline for Bottom-Up Panoptic Segmentationno summary yetcs-cv1911.10194Google Research42 citesNov 22, 2019
- NovStructEdit: Learning Structural Shape Variationsno summary yetcs-cv1911.11098Adobe2 citesNov 25, 2019
- NovRevisiting Image Aesthetic Assessment via Self-Supervised Feature Learningno summary yetcs-cv1911.11419Tencent7 citesNov 26, 2019
- NovViewAL: Active Learning with Viewpoint Entropy for Semantic Segmentationno summary yetcs-cv1911.11789Google Research10 citesNov 26, 2019
- NovDocument Structure Extraction using Prior based High Resolution Hierarchical Semantic Segmentationno summary yetcs-cv1911.12170Adobe1 citesNov 27, 2019
- NovAutoRemover: Automatic Object Removal for Autonomous Driving Videosno summary yetcs-cv1911.12588Baidu2 citesNov 28, 2019
- NovAttributional Robustness Training using Input-Gradient Spatial Alignmentno summary yetcs-cv1911.13073Adobe5 citesNov 29, 2019
- NovDIST: Rendering Deep Implicit Signed Distance Function with Differentiable Sphere Tracingno summary yetcs-cv1911.13225Microsoft Research19 citesNov 29, 2019
- NovDomain-invariant Stereo Matching Networksno summary yetcs-cv1911.13287Baidu6 citesNov 29, 2019
- NovWhat's Hidden in a Randomly Weighted Neural Network?no summary yetcs-cv1911.13299AllenAI26 citesNov 29, 2019
- OctPrivacy-preserving Federated Brain Tumour Segmentationno summary yetcs-cv1910.00962NVIDIA63 citesOct 2, 2019
- OctWeakly supervised segmentation from extreme pointsno summary yetcs-cv1910.01236NVIDIA20 citesOct 2, 2019
- OctCLEVRER: CoLlision Events for Video REpresentation and Reasoningno summary yetcs-cv1910.01442Google Research70 citesOct 3, 2019
- Oct3D Neighborhood Convolution: Learning Depth-Aware Features for RGB-D and RGB Semantic Segmentationno summary yetcs-cv1910.01460Google Research2 citesOct 3, 2019
- OctNeural Puppet: Generative Layered Cartoon Charactersno summary yetcs-cv1910.02060Adobe1 citesOct 4, 2019
- OctTo React or not to React: End-to-End Visual Pose Forecasting for Personalized Avatar during Dyadic Conversationsno summary yetcs-cv1910.02181Meta / FAIR1 citesOct 5, 2019
- OctSelf-supervised Feature Learning for 3D Medical Images by Playing a Rubik's Cubeno summary yetcs-cv1910.02241Tencent15 citesOct 5, 2019
- OctDexPilot: Vision Based Teleoperation of Dexterous Robotic Hand-Arm Systemno summary yetcs-cv1910.03135NVIDIA10 citesOct 7, 2019
- OctVisual Indeterminacy in GAN Artno summary yetcs-cv1910.04639Adobe30 citesOct 10, 2019
- OctRosetta: Large scale system for text detection and recognition in imagesno summary yetcs-cv1910.05085Meta / FAIR335 citesOct 11, 2019
- OctOne-Shot Neural Architecture Search via Self-Evaluated Template Networkno summary yetcs-cv1910.05733Baidu207 citesOct 13, 2019
- OctBuilding Damage Detection in Satellite Imagery Using Convolutional Neural Networksno summary yetcs-cv1910.06444Google Research116 citesOct 14, 2019
- OctEnd-to-End Adversarial Shape Learning for Abdomen Organ Deep Segmentationno summary yetcs-cv1910.06474NVIDIA12 citesOct 15, 2019
- OctGenerative Modeling for Small-Data Object Detectionno summary yetcs-cv1910.07169Google Research2 citesOct 16, 2019
- OctA Dataset of Multi-Illumination Images in the Wildno summary yetcs-cv1910.08131Adobe1 citesOct 17, 2019
- OctDeep Parametric Indoor Lighting Estimationno summary yetcs-cv1910.08812Adobe11 citesOct 19, 2019
- OctExploring Simple and Transferable Recognition-Aware Image Processingno summary yetcs-cv1910.09185Adobe9 citesOct 21, 2019
- OctAggregation Signature for Small Object Trackingno summary yetcs-cv1910.10859Baidu30 citesOct 24, 2019
- OctControllable Attention for Structured Layered Video Decompositionno summary yetcs-cv1910.11306DeepMind0 citesOct 24, 2019
- OctHandheld Mobile Photography in Very Low Lightno summary yetcs-cv1910.11336Google Research107 citesOct 24, 2019
- OctSelf-supervised Learning of Detailed 3D Face Reconstructionno summary yetcs-cv1910.11791Tencent62 citesOct 25, 2019
- OctLearning to Track Any Objectno summary yetcs-cv1910.11844Google Research4 citesOct 25, 2019
- OctImage-Based Place Recognition on Bucolic Environment Across Seasons From Semantic Edge Descriptionno summary yetcs-cv1910.12468Google Research32 citesOct 28, 2019
- OctClassification Calibration for Long-tail Instance Segmentationno summary yetcs-cv1910.13081Salesforce9 citesOct 29, 2019
- OctSemantic Conditioned Dynamic Modulation for Temporal Sentence Grounding in Videosno summary yetcs-cv1910.14303Tencent0 citesOct 31, 2019
- OctMaking an Invisibility Cloak: Real World Adversarial Attacks on Object Detectorsno summary yetcs-cv1910.14667Meta / FAIR14 citesOct 31, 2019
- SepDomain Randomization and Pyramid Consistency: Simulation-to-Real Generalization without Accessing Target Domain Datano summary yetcs-cv1909.00889Google Research20 citesSep 2, 2019
- SepEleAtt-RNN: Adding Attentiveness to Neurons in Recurrent Neural Networksno summary yetcs-cv1909.01939Alibaba107 citesSep 3, 2019
- SepBeyond Photo Realism for Domain Adaptation from Synthetic Datano summary yetcs-cv1909.01960Google Research1 citesSep 4, 2019
- SepLarge-scale Tag-based Font Retrieval with Generative Feature Learningno summary yetcs-cv1909.02072Adobe1 citesSep 4, 2019
- SepDeep Visual Template-Free Form Parsingno summary yetcs-cv1909.02576Adobe1 citesSep 5, 2019
- SepNon-discriminative data or weak model? On the relative importance of data and model resolutionno summary yetcs-cv1909.03205Google Research6 citesSep 7, 2019
- SepOpen Compound Domain Adaptationno summary yetcs-cv1909.03403Google Research2 citesSep 8, 2019
- SepFreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape from Single RGB Imagesno summary yetcs-cv1909.04349Adobe46 citesSep 10, 2019
- SepThe Mapillary Traffic Sign Dataset for Detection and Classification on a Global Scaleno summary yetcs-cv1909.04422Meta / FAIR5 citesSep 10, 2019
- SepTemporally Grounding Language Queries in Videos by Contextual Boundary-aware Predictionno summary yetcs-cv1909.05010Tencent20 citesSep 11, 2019
- SepCvxNet: Learnable Convex Decompositionno summary yetcs-cv1909.05736Google Research7 citesSep 12, 2019
- SepVisuomotor Understanding for Representation Learning of Driving Scenesno summary yetcs-cv1909.06979Microsoft Research6 citesSep 16, 2019
- SepLearning Visuomotor Policies for Aerial Navigation Using Cross-Modal Representationsno summary yetcs-cv1909.06993Microsoft Research7 citesSep 16, 2019
- SepMultimodal Multitask Representation Learning for Pathology Biobank Metadata Predictionno summary yetcs-cv1909.07846Google Research17 citesSep 17, 2019
- SepMulti-mapping Image-to-Image Translation via Learning Disentanglementno summary yetcs-cv1909.07877Tencent45 citesSep 17, 2019
- SepA Data-Center FPGA Acceleration Platform for Convolutional Neural Networksno summary yetcs-cv1909.07973Tencent21 citesSep 17, 2019
- SepLarge-scale representation learning from visually grounded untranscribed speechno summary yetcs-cv1909.08782Google Research2 citesSep 19, 2019
- SepInteractive Sketch & Fill: Multiclass Sketch-to-Image Translationno summary yetcs-cv1909.11081Adobe16 citesSep 24, 2019
- SepAttention Convolutional Binary Neural Tree for Fine-Grained Visual Categorizationno summary yetcs-cv1909.11378Tencent14 citesSep 25, 2019
- SepWiderPerson: A Diverse Dataset for Dense Pedestrian Detection in the Wildno summary yetcs-cv1909.12118Baidu2 citesSep 25, 2019
- SepImplicit Semantic Data Augmentation for Deep Networksno summary yetcs-cv1909.12220Baidu60 citesSep 26, 2019
- SepTowards Object Detection from Motionno summary yetcs-cv1909.12950Google Research1 citesSep 17, 2019
- SepRandAugment: Practical automated data augmentation with a reduced search spaceno summary yetcs-cv1909.13719Google Research281 citesSep 30, 2019
- SepPlace Deduplication with Embeddingsno summary yetcs-cv1910.04861Meta / FAIR0 citesSep 29, 2019
- AugCentral Similarity Quantization for Efficient Image and Video Retrievalno summary yetcs-cv1908.00347Tencent19 citesAug 1, 2019
- AugMoulding Humans: Non-parametric 3D Human Shape Estimation from Single Imagesno summary yetcs-cv1908.00439Google Research10 citesAug 1, 2019
- AugA Unified Point-Based Framework for 3D Segmentationno summary yetcs-cv1908.00478Amazon0 citesAug 1, 2019
- AugLearning Local Feature Descriptor with Motion Attribute for Vision-based Localizationno summary yetcs-cv1908.01180Alibaba0 citesAug 3, 2019
- AugReal-Time Global Illumination Decomposition of Videosno summary yetcs-cv1908.01961Google Research23 citesAug 6, 2019
- AugSemi-Supervised Adversarial Monocular Depth Estimationno summary yetcs-cv1908.02126Microsoft Research5 citesAug 6, 2019
- AugProgressive Transfer Learningno summary yetcs-cv1908.02492Alibaba1 citesAug 7, 2019
- AugUnsupervised Feature Learning in Remote Sensingno summary yetcs-cv1908.02877NVIDIA3 citesAug 7, 2019
- AugDeep Density-aware Count Regressorno summary yetcs-cv1908.03314Baidu10 citesAug 9, 2019
- AugA Distraction Score for Watermarksno summary yetcs-cv1908.03651Google Research0 citesAug 9, 2019
- AugRecent Advances in Deep Learning for Object Detectionno summary yetcs-cv1908.03673Salesforce25 citesAug 10, 2019
- AugConditional Generative Adversarial Networks for Data Augmentation and Adaptation in Remotely Sensed Imageryno summary yetcs-cv1908.03809NVIDIA22 citesAug 10, 2019
- AugIoU Loss for 2D/3D Object Detectionno summary yetcs-cv1908.03851Baidu16 citesAug 11, 2019
- AugSentence Specified Dynamic Video Thumbnail Generationno summary yetcs-cv1908.04052Tencent32 citesAug 12, 2019
- AugLearning elementary structures for 3D shape generation and matchingno summary yetcs-cv1908.04725Adobe15 citesAug 13, 2019
- AugA Single-Shot Arbitrarily-Shaped Text Detector based on Context Attended Multi-Task Learningno summary yetcs-cv1908.05498Baidu81 citesAug 15, 2019
- AugRIO: 3D Object Instance Re-Localization in Changing Indoor Environmentsno summary yetcs-cv1908.06109Google Research9 citesAug 16, 2019
- AugNeural Re-Simulation for Generating Bounces in Single Imagesno summary yetcs-cv1908.06217Adobe1 citesAug 17, 2019
- AugConvolutional Neural Network with Median Layers for Denoising Salt-and-Pepper Contaminationsno summary yetcs-cv1908.06452Microsoft Research7 citesAug 18, 2019
- AugBoundless: Generative Adversarial Networks for Image Extensionno summary yetcs-cv1908.07007Google Research34 citesAug 19, 2019
- AugAction recognition with spatial-temporal discriminative filter banksno summary yetcs-cv1908.07625Amazon10 citesAug 20, 2019
- AugSaccader: Improving Accuracy of Hard Attention Models for Visionno summary yetcs-cv1908.07644Google Research11 citesAug 20, 2019
- AugRBCN: Rectified Binary Convolutional Networks for Enhancing the Performance of 1-bit DCNNsno summary yetcs-cv1908.07748Baidu5 citesAug 21, 2019
- AugDeep High-Resolution Representation Learning for Visual Recognitionno summary yetcs-cv1908.07919Microsoft Research354 citesAug 20, 2019
- AugEnd-to-End Boundary Aware Networks for Medical Image Segmentationno summary yetcs-cv1908.08071NVIDIA4 citesAug 21, 2019
- AugSequential Latent Spaces for Modeling the Intention During Diverse Image Captioningno summary yetcs-cv1908.08529Meta / FAIR3 citesAug 22, 2019
- AugOnion-Peel Networks for Deep Video Completionno summary yetcs-cv1908.08718Adobe9 citesAug 23, 2019
- AugObject-Driven Multi-Layer Scene Decomposition From a Single Imageno summary yetcs-cv1908.09521Google Research5 citesAug 26, 2019
- AugCASIA-SURF: A Large-scale Multi-modal Benchmark for Face Anti-spoofingno summary yetcs-cv1908.10654Baidu17 citesAug 28, 2019
- AugBoundary-Aware Feature Propagation for Scene Segmentationno summary yetcs-cv1909.00179Alibaba28 citesAug 31, 2019
- AugPush for Quantization: Deep Fisher Hashingno summary yetcs-cv1909.00206Tencent2 citesAug 31, 2019
- AugWSLLN: Weakly Supervised Natural Language Localization Networksno summary yetcs-cv1909.00239Salesforce5 citesAug 31, 2019
- JulGoing Deeper with Lean Point Networksno summary yetcs-cv1907.00960Adobe1 citesJul 1, 2019
- JulSelf-supervised Learning of Interpretable Keypoints from Unlabelled Videosno summary yetcs-cv1907.02055DeepMind57 citesJul 3, 2019
- JulThe Indirect Convolution Algorithmno summary yetcs-cv1907.02129Google Research28 citesJul 3, 2019
- JulSim2real transfer learning for 3D human pose estimation: motion to the rescueno summary yetcs-cv1907.02499Google Research71 citesJul 4, 2019
- JulLarge Scale Adversarial Representation Learningno summary yetcs-cv1907.02544Google Research260 citesJul 4, 2019
- JulACNe: Attentive Context Normalization for Robust Permutation-Equivariant Learningno summary yetcs-cv1907.02545Google Research10 citesJul 4, 2019
- JulFast Universal Style Transfer for Artistic and Photorealistic Renderingno summary yetcs-cv1907.03118Baidu4 citesJul 6, 2019
- JulRevisiting Metric Learning for Few-Shot Image Classificationno summary yetcs-cv1907.03123Tencent5 citesJul 6, 2019
- JulUnsupervised cycle-consistent deformation for shape matchingno summary yetcs-cv1907.03165Adobe1 citesJul 6, 2019
- JulELF: Embedded Localisation of Features in pre-trained CNNno summary yetcs-cv1907.03261Google Research2 citesJul 7, 2019
- JulAttentive CT Lesion Detection Using Deep Pyramid Inference with Multi-Scale Boosterno summary yetcs-cv1907.03958Tencent2 citesJul 9, 2019
- JulImproving Deep Lesion Detection Using 3D Contextual and Spatial Attentionno summary yetcs-cv1907.04052NVIDIA5 citesJul 9, 2019
- JulBlazeFace: Sub-millisecond Neural Face Detection on Mobile GPUsno summary yetcs-cv1907.05047Google Research249 citesJul 11, 2019
- JulCross-Domain Complementary Learning Using Pose for Multi-Person Part Segmentationno summary yetcs-cv1907.05193Microsoft Research73 citesJul 11, 2019
- JulAdversarial Video Generation on Complex Datasetsno summary yetcs-cv1907.06571Google Research148 citesJul 15, 2019
- JulReal-time Facial Surface Geometry from Monocular Video on Mobile GPUsno summary yetcs-cv1907.06724Google Research58 citesJul 15, 2019
- JulReal-time Hair Segmentation and Recoloring on Mobile GPUsno summary yetcs-cv1907.06740Google Research8 citesJul 15, 2019
- JulRethinking RGB-D Salient Object Detection: Models, Data Sets, and Large-Scale Benchmarksno summary yetcs-cv1907.06781Google Research688 citesJul 15, 2019
- Jul2nd Place Solution to the GQA Challenge 2019no summary yetcs-cv1907.06794Amazon6 citesJul 16, 2019
- JulInstant Motion Tracking and Its Applications to Augmented Realityno summary yetcs-cv1907.06796Google Research6 citesJul 16, 2019
- JulDomain-Specific Priors and Meta Learning for Few-Shot First-Person Action Recognitionno summary yetcs-cv1907.09382Microsoft Research30 citesJul 22, 2019
- JulInformation-Bottleneck Approach to Salient Region Discoveryno summary yetcs-cv1907.09578Google Research12 citesJul 22, 2019
- JulMixConv: Mixed Depthwise Convolutional Kernelsno summary yetcs-cv1907.09595Google Research52 citesJul 22, 2019
- JulReal-Time Correlation Tracking via Joint Model Compression and Transferno summary yetcs-cv1907.09831Tencent54 citesJul 23, 2019
- JulDR Loss: Improving Object Detection by Distributional Rankingno summary yetcs-cv1907.10156Alibaba9 citesJul 23, 2019
- JulMixed-Supervised Dual-Network for Medical Image Segmentationno summary yetcs-cv1907.10209Microsoft Research4 citesJul 24, 2019
- JulDual Grid Net: hand mesh vertex regression from single depth mapsno summary yetcs-cv1907.10695Meta / FAIR2 citesJul 24, 2019
- JulROAM: Recurrently Optimizing Tracking Modelno summary yetcs-cv1907.12006Tencent3 citesJul 28, 2019
- JulChaLearn Looking at People: IsoGD and ConGD Large-scale RGB-D Gesture Recognitionno summary yetcs-cv1907.12193Baidu36 citesJul 29, 2019
- JulFSS-1000: A 1000-Class Dataset for Few-Shot Segmentationno summary yetcs-cv1907.12347Tencent234 citesJul 29, 2019
- JulConsensus Feature Network for Scene Parsingno summary yetcs-cv1907.12411Baidu3 citesJul 29, 2019
- JulLEAF-QA: Locate, Encode & Attend for Figure Question Answeringno summary yetcs-cv1907.12861Adobe12 citesJul 30, 2019
- JulUnsupervised Separation of Dynamics from Pixelsno summary yetcs-cv1907.12906DeepMind3 citesJul 20, 2019
- JulWeakly Supervised Body Part Segmentation with Pose based Part Priorsno summary yetcs-cv1907.13051Google Research2 citesJul 30, 2019
- JulDeblurring Face Images using Uncertainty Guided Multi-Stream Semantic Networksno summary yetcs-cv1907.13106Adobe73 citesJul 30, 2019
- JulMulti-Agent Reinforcement Learning Based Frame Sampling for Effective Untrimmed Video Recognitionno summary yetcs-cv1907.13369Baidu27 citesJul 31, 2019
- JunLearning to Generate Grounded Visual Captions without Localization Supervisionno summary yetcs-cv1906.00283Meta / FAIR5 citesJun 1, 2019
- JunHierarchical Video Frame Sequence Representation with Deep Convolutional Graph Networkno summary yetcs-cv1906.00377Alibaba28 citesJun 2, 2019
- JunReconstruct and Represent Video Contents for Captioning via Reinforcement Learningno summary yetcs-cv1906.01452Tencent5 citesJun 3, 2019
- JunText-based Editing of Talking-head Videono summary yetcs-cv1906.01524Adobe5 citesJun 4, 2019
- JunNatural Vocabulary Emerges from Free-Form Annotationsno summary yetcs-cv1906.01542Google Research2 citesJun 4, 2019
- JunStyleNAS: An Empirical Study of Neural Architecture Search to Uncover Surprisingly Fast End-to-End Universal Style Transfer Networksno summary yetcs-cv1906.02470Baidu3 citesJun 6, 2019
- JunScaling Autoregressive Video Modelsno summary yetcs-cv1906.02634Google Research14 citesJun 6, 2019
- JunXRAI: Better Attributions Through Regionsno summary yetcs-cv1906.02825Google Research7 citesJun 6, 2019
- JunExtracting Visual Knowledge from the Internet: Making Sense of Image Datano summary yetcs-cv1906.03219Alibaba0 citesJun 7, 2019
- JunEvolving Losses for Unlabeled Video Representation Learningno summary yetcs-cv1906.03248Google Research4 citesJun 7, 2019
- JunHowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clipsno summary yetcs-cv1906.03327DeepMind117 citesJun 7, 2019
- JunFast Spatially-Varying Indoor Lighting Estimationno summary yetcs-cv1906.03799Adobe2 citesJun 10, 2019
- Jun2nd Place and 2nd Place Solution to Kaggle Landmark Recognition andRetrieval Competition 2019no summary yetcs-cv1906.03990Baidu6 citesJun 10, 2019
- JunScale Invariant Fully Convolutional Network: Detecting Hands Efficientlyno summary yetcs-cv1906.04634Tencent4 citesJun 11, 2019
- JunThe Herbarium Challenge 2019 Datasetno summary yetcs-cv1906.05372Google Research26 citesJun 12, 2019
- JunAssisted Excitation of Activations: A Learning Technique to Improve Object Detectorsno summary yetcs-cv1906.05388AllenAI5 citesJun 12, 2019
- JunUnsupervised Monocular Depth and Ego-motion Learning with Structure and Semanticsno summary yetcs-cv1906.05717Google Research13 citesJun 12, 2019
- JunVisual Wake Words Datasetno summary yetcs-cv1906.05721Google Research84 citesJun 12, 2019
- JunContrastive Multiview Codingno summary yetcs-cv1906.05849Google Research571 citesJun 13, 2019
- JunStand-Alone Self-Attention in Vision Modelsno summary yetcs-cv1906.05909Google Research222 citesJun 13, 2019
- JunUnsupervised Video Interpolation Using Cycle Consistencyno summary yetcs-cv1906.05928Google Research11 citesJun 13, 2019
- JunImage Counterfactual Sensitivity Analysis for Detecting Unintended Biasno summary yetcs-cv1906.06439Google Research19 citesJun 14, 2019
- JunFloors are Flat: Leveraging Semantics for Real-Time Surface Normal Predictionno summary yetcs-cv1906.06792Google Research2 citesJun 16, 2019
- JunPanoptic Image Annotation with a Collaborative Assistantno summary yetcs-cv1906.06798Google Research1 citesJun 17, 2019
- JunEnlightenGAN: Deep Light Enhancement without Paired Supervisionno summary yetcs-cv1906.06972Microsoft Research205 citesJun 17, 2019
- JunDeepView: View Synthesis with Learned Gradient Descentno summary yetcs-cv1906.07316Google Research21 citesJun 18, 2019
- JunUnsupervised Learning of Object Structure and Dynamics from Videosno summary yetcs-cv1906.07889Adobe19 citesJun 19, 2019
- JunDeep RGB-D Canonical Correlation Analysis For Sparse Depth Completionno summary yetcs-cv1906.08967Adobe10 citesJun 21, 2019
- JunRUBi: Reducing Unimodal Biases in Visual Question Answeringno summary yetcs-cv1906.10169Meta / FAIR205 citesJun 24, 2019
- JunLearning Data Augmentation Strategies for Object Detectionno summary yetcs-cv1906.11172Google Research152 citesJun 26, 2019
- JunEmergence of Exploratory Look-Around Behaviors through Active Observation Completionno summary yetcs-cv1906.11407Meta / FAIR37 citesJun 27, 2019
- JunUnsupervised Learning of Object Keypoints for Perception and Controlno summary yetcs-cv1906.11883Google Research54 citesJun 19, 2019
- JunLarge-scale, real-time visual-inertial localization revisitedno summary yetcs-cv1907.00338Google Research2 citesJun 30, 2019
- MayPushing the Boundaries of View Extrapolation with Multiplane Imagesno summary yetcs-cv1905.00413Google Research10 citesMay 1, 2019
- MayDPSNet: End-to-end Deep Plane Sweep Stereono summary yetcs-cv1905.00538Microsoft Research84 citesMay 2, 2019
- MaySinGAN: Learning a Generative Model from a Single Natural Imageno summary yetcs-cv1905.01164Google Research97 citesMay 2, 2019
- MaySteadiface: Real-Time Face-Centric Stabilization on Mobile Phonesno summary yetcs-cv1905.01382Google Research1 citesMay 3, 2019
- MayPasteGAN: A Semi-Parametric Method to Generate Image from Scene Graphno summary yetcs-cv1905.01608Microsoft Research27 citesMay 5, 2019
- MayDeep Video Inpaintingno summary yetcs-cv1905.01639Adobe5 citesMay 5, 2019
- MayNostalgin: Extracting 3D City Models from Historical Image Datano summary yetcs-cv1905.01772Google Research4 citesMay 6, 2019
- MaySearching for MobileNetV3no summary yetcs-cv1905.02244Google Research431 citesMay 6, 2019
- MayAdapting Image Super-Resolution State-of-the-arts and Learning Multi-model Ensemble for Video Super-Resolutionno summary yetcs-cv1905.02462Baidu1 citesMay 7, 2019
- MayContrastive Learning for Lifted Networksno summary yetcs-cv1905.02507Microsoft Research1 citesMay 7, 2019
- MayHigh Frequency Residual Learning for Multi-Scale Image Classificationno summary yetcs-cv1905.02649Microsoft Research2 citesMay 7, 2019
- MayInverse Rendering for Complex Indoor Scenes: Shape, Spatially-Varying Lighting and SVBRDF from a Single Imageno summary yetcs-cv1905.02722Adobe16 citesMay 7, 2019
- MayHandheld Multi-Frame Super-Resolutionno summary yetcs-cv1905.03277Google Research213 citesMay 8, 2019
- MayGrand Challenge of 106-Point Facial Landmark Localizationno summary yetcs-cv1905.03469Baidu5 citesMay 9, 2019
- MayS4L: Self-Supervised Semi-Supervised Learningno summary yetcs-cv1905.03670Google Research187 citesMay 9, 2019
- MayDeep Sky Modeling for Single Image Outdoor Lighting Estimationno summary yetcs-cv1905.03897Adobe6 citesMay 10, 2019
- MayDeepICP: An End-to-End Deep Neural Network for 3D Point Cloud Registrationno summary yetcs-cv1905.04153Baidu279 citesMay 10, 2019
- MayExploiting temporal context for 3D human pose estimation in the wildno summary yetcs-cv1905.04266DeepMind7 citesMay 10, 2019
- MayNTU RGB+D 120: A Large-Scale Benchmark for 3D Human Activity Understandingno summary yetcs-cv1905.04757Alibaba1,721 citesMay 12, 2019
- MayListwise View Ranking for Image Croppingno summary yetcs-cv1905.05352Tencent20 citesMay 14, 2019
- MayReconstruction-Aware Imaging System Ranking by use of a Sparsity-Driven Numerical Observer Enabled by Variational Bayesian Inferenceno summary yetcs-cv1905.05820Google Research4 citesMay 14, 2019
- MayRelaxed 2-D Principal Component Analysis by $L_p$ Norm for Face Recognitionno summary yetcs-cv1905.06458Baidu1 citesMay 15, 2019
- MayHow is Gaze Influenced by Image Transformations? Dataset and Modelno summary yetcs-cv1905.06803Baidu78 citesMay 16, 2019
- MayLearning to Reconstruct 3D Manhattan Wireframes from a Single Imageno summary yetcs-cv1905.07482Adobe7 citesMay 17, 2019
- MaySplitNet: Sim2Sim and Task2Task Transfer for Embodied Visual Navigationno summary yetcs-cv1905.07512Meta / FAIR11 citesMay 18, 2019
- MayWhich Tasks Should Be Learned Together in Multi-task Learning?no summary yetcs-cv1905.07553Google Research68 citesMay 18, 2019
- MayBoundary Learning by Using Weighted Propagation in Convolution Networkno summary yetcs-cv1905.09226Tencent16 citesMay 22, 2019
- MayData-Efficient Image Recognition with Contrastive Predictive Codingno summary yetcs-cv1905.09272Google Research937 citesMay 22, 2019
- MaySpeech2Face: Learning the Face Behind a Voiceno summary yetcs-cv1905.09773Google Research13 citesMay 23, 2019
- MayEnsembleNet: End-to-End Optimization of Multi-headed Modelsno summary yetcs-cv1905.09979Google Research14 citesMay 24, 2019
- MayFrom Here to There: Video Inbetweening Using Direct 3D Convolutionsno summary yetcs-cv1905.10240Google Research20 citesMay 24, 2019
- MayRobust Unsupervised Flexible Auto-weighted Local-Coordinate Concept Factorization for Image Clusteringno summary yetcs-cv1905.10564Adobe4 citesMay 25, 2019
- MayEfficient Object Annotation via Speaking and Pointingno summary yetcs-cv1905.10576Google Research0 citesMay 25, 2019
- MayDISN: Deep Implicit Surface Network for High-quality Single-view 3D Reconstructionno summary yetcs-cv1905.10711Adobe240 citesMay 26, 2019
- MayFooling Detection Alone is Not Enough: First Adversarial Attack against Multiple Object Trackingno summary yetcs-cv1905.11026Baidu56 citesMay 27, 2019
- MaySemantic Fisher Scores for Task Transfer: Using Objects to Classify Scenesno summary yetcs-cv1905.11539Microsoft Research2 citesMay 27, 2019
- MayContextual Translation Embedding for Visual Relationship Detection and Scene Graph Generationno summary yetcs-cv1905.11624NVIDIA87 citesMay 28, 2019
- MayHallucinating Optical Flow Features for Video Classificationno summary yetcs-cv1905.11799Tencent3 citesMay 28, 2019
- MayCerberus: A Multi-headed Derendererno summary yetcs-cv1905.11940Google Research6 citesMay 28, 2019
- MayVolumetric Capture of Humans with a Single RGBD Camera via Semi-Parametric Learningno summary yetcs-cv1905.12162Google Research0 citesMay 29, 2019
- MayAlign-and-Attend Network for Globally and Locally Coherent Video Inpaintingno summary yetcs-cv1905.13066Adobe0 citesMay 30, 2019
- MayA Hierarchical Probabilistic U-Net for Modeling Multi-Scale Ambiguitiesno summary yetcs-cv1905.13077Google Research27 citesMay 30, 2019
- MayAssembleNet: Searching for Multi-Stream Neural Connectivity in Video Architecturesno summary yetcs-cv1905.13209Google Research27 citesMay 30, 2019
- MayMultitask Text-to-Visual Embedding with Titles and Clickthrough Datano summary yetcs-cv1905.13339Adobe2 citesMay 30, 2019
- MayOK-VQA: A Visual Question Answering Benchmark Requiring External Knowledgeno summary yetcs-cv1906.00067AllenAI10 citesMay 31, 2019
- AprVideo Object Segmentation using Space-Time Memory Networksno summary yetcs-cv1904.00607Adobe73 citesApr 1, 2019
- AprStandardized Assessment of Automatic Segmentation of White Matter Hyperintensities and Results of the WMH Segmentation Challengeno summary yetcs-cv1904.00682IBM Research301 citesApr 1, 2019
- AprCreativity Inspired Zero-Shot Learningno summary yetcs-cv1904.01109Baidu7 citesApr 1, 2019
- AprDeepLight: Learning Illumination for Unconstrained Mobile Mixed Realityno summary yetcs-cv1904.01175Google Research4 citesApr 2, 2019
- Apr3DRegNet: A Deep Neural Network for 3D Point Registrationno summary yetcs-cv1904.01701Google Research19 citesApr 2, 2019
- AprA Learned Representation for Scalable Vector Graphicsno summary yetcs-cv1904.02632Google Research15 citesApr 4, 2019
- AprEmbodied Question Answering in Photorealistic Environments with Point Cloud Perceptionno summary yetcs-cv1904.03461Meta / FAIR13 citesApr 6, 2019
- AprMeasuring Human Perception to Improve Handwritten Document Transcriptionno summary yetcs-cv1904.03734Adobe2 citesApr 7, 2019
- AprMeta-Learning with Differentiable Convex Optimizationno summary yetcs-cv1904.03758Amazon96 citesApr 7, 2019
- AprRevisiting EmbodiedQA: A Simple Baseline and Beyondno summary yetcs-cv1904.04166Google Research25 citesApr 8, 2019
- AprNeural Rerendering in the Wildno summary yetcs-cv1904.04290Google Research9 citesApr 8, 2019
- AprLabel Propagation for Deep Semi-supervised Learningno summary yetcs-cv1904.04717Google Research25 citesApr 9, 2019
- AprCondConv: Conditionally Parameterized Convolutions for Efficient Inferenceno summary yetcs-cv1904.04971Google Research284 citesApr 10, 2019
- AprDepth from Videos in the Wild: Unsupervised Monocular Depth Learning from Unknown Camerasno summary yetcs-cv1904.04998Google Research44 citesApr 10, 2019
- AprSemi-Supervised Graph Classification: A Hierarchical Graph Perspectiveno summary yetcs-cv1904.05003Tencent1 citesApr 10, 2019
- AprH+O: Unified Egocentric Recognition of 3D Hand-Object Poses and Interactionsno summary yetcs-cv1904.05349Microsoft Research6 citesApr 10, 2019
- AprLearning to Generate Synthetic Data via Compositingno summary yetcs-cv1904.05475Amazon6 citesApr 10, 2019
- AprPredicting Progression of Age-related Macular Degeneration from Fundus Images using Deep Learningno summary yetcs-cv1904.05478Google Research10 citesApr 10, 2019
- AprLearning Single Camera Depth Estimation using Dual-Pixelsno summary yetcs-cv1904.05822Google Research16 citesApr 11, 2019
- AprTwo Body Problem: Collaborative Visual Task Completionno summary yetcs-cv1904.05879AllenAI11 citesApr 11, 2019
- AprSynthetic Examples Improve Generalization for Rare Classesno summary yetcs-cv1904.05916Microsoft Research15 citesApr 11, 2019
- AprEvalNorm: Estimating Batch Normalization Statistics for Evaluationno summary yetcs-cv1904.06031Google Research7 citesApr 12, 2019
- AprDetecting Anemia from Retinal Fundus Imagesno summary yetcs-cv1904.06435Google Research223 citesApr 12, 2019
- AprLearning Shape Templates with Structured Implicit Functionsno summary yetcs-cv1904.06447Google Research44 citesApr 12, 2019
- AprBounce and Learn: Modeling Scene Dynamics with Real-World Bouncesno summary yetcs-cv1904.06827Adobe10 citesApr 15, 2019
- AprNAS-FPN: Learning Scalable Feature Pyramid Architecture for Object Detectionno summary yetcs-cv1904.07392Google Research355 citesApr 16, 2019
- AprShared Predictive Cross-Modal Deep Quantizationno summary yetcs-cv1904.07488Tencent0 citesApr 16, 2019
- AprTemporal Cycle-Consistency Learningno summary yetcs-cv1904.07846DeepMind16 citesApr 16, 2019
- AprAutomated Design of Deep Learning Methods for Biomedical Image Segmentationno summary yetcs-cv1904.08128DeepMind8,401 citesApr 17, 2019
- AprAutomated Segmentation of Pulmonary Lobes using Coordination-Guided Deep Neural Networksno summary yetcs-cv1904.09106Alibaba27 citesApr 19, 2019
- AprA Scalable Handwritten Text Recognition Systemno summary yetcs-cv1904.09150Google Research12 citesApr 19, 2019
- AprFashion++: Minimal Edits for Outfit Improvementno summary yetcs-cv1904.09261Meta / FAIR9 citesApr 19, 2019
- AprFast User-Guided Video Object Segmentation by Interaction-and-Propagation Networksno summary yetcs-cv1904.09791Adobe4 citesApr 22, 2019
- AprAttention Augmented Convolutional Networksno summary yetcs-cv1904.09925Google Research198 citesApr 22, 2019
- AprUsing Videos to Evaluate Image Model Robustnessno summary yetcs-cv1904.10076Google Research31 citesApr 22, 2019
- AprPath-Restore: Learning Network Path Selection for Image Restorationno summary yetcs-cv1904.10343Tencent107 citesApr 23, 2019
- AprNeural Collaborative Subspace Clusteringno summary yetcs-cv1904.10596Tencent13 citesApr 24, 2019
- AprWeb Stereo Video Supervision for Depth Prediction from Dynamic Scenesno summary yetcs-cv1904.11112Adobe6 citesApr 25, 2019
- AprMaking Convolutional Networks Shift-Invariant Againno summary yetcs-cv1904.11486Adobe104 citesApr 25, 2019
- AprThe Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervisionno summary yetcs-cv1904.12584Google Research314 citesApr 26, 2019
- AprPredicting How to Distribute Work Between Algorithms and Humans to Segment an Image Batchno summary yetcs-cv1905.00060Meta / FAIR2 citesApr 30, 2019
- MarExtreme Channel Prior Embedded Network for Dynamic Scene Deblurringno summary yetcs-cv1903.00763Alibaba99 citesMar 2, 2019
- MarGround Plane based Absolute Scale Estimation for Monocular Visual Odometryno summary yetcs-cv1903.00912Baidu6 citesMar 3, 2019
- MarVideoFlow: A Conditional Flow-Based Model for Stochastic Video Generationno summary yetcs-cv1903.01434Google Research76 citesMar 4, 2019
- MarTableBank: A Benchmark Dataset for Table Detection and Recognitionno summary yetcs-cv1903.01949Microsoft Research60 citesMar 5, 2019
- Mar3DN: 3D Deformation Networkno summary yetcs-cv1903.03322Adobe10 citesMar 8, 2019
- MarSliced Wasserstein Discrepancy for Unsupervised Domain Adaptationno summary yetcs-cv1903.04064Apple66 citesMar 10, 2019
- MarHierarchical Autoregressive Image Models with Auxiliary Decodersno summary yetcs-cv1903.04933Google Research26 citesMar 6, 2019
- MarUnsupervised Discovery of Parts, Structure, and Dynamicsno summary yetcs-cv1903.05136Google Research25 citesMar 12, 2019
- MarPrivacy Preserving Image-Based Localizationno summary yetcs-cv1903.05572Microsoft Research4 citesMar 13, 2019
- MarMSG-GAN: Multi-Scale Gradients for Generative Adversarial Networksno summary yetcs-cv1903.06048Adobe24 citesMar 14, 2019
- MarLive Reconstruction of Large-Scale Dynamic Outdoor Worldsno summary yetcs-cv1903.06708Microsoft Research1 citesMar 15, 2019
- MarA Cross-Season Correspondence Dataset for Robust Semantic Segmentationno summary yetcs-cv1903.06916Microsoft Research11 citesMar 16, 2019
- MarBilinear Representation for Language-based Image Editing Using Conditional Generative Adversarial Networksno summary yetcs-cv1903.07499Alibaba26 citesMar 18, 2019
- MarUnderstanding the Limitations of CNN-based Absolute Camera Pose Regressionno summary yetcs-cv1903.07504Microsoft Research21 citesMar 18, 2019
- MarCross-task weakly supervised learning from instructional videosno summary yetcs-cv1903.08225DeepMind8 citesMar 19, 2019
- MarLooking Fast and Slow: Memory-Guided Mobile Video Object Detectionno summary yetcs-cv1903.10172Google Research30 citesMar 25, 2019
- MarDeepRED: Deep Image Prior Powered by REDno summary yetcs-cv1903.10176Google Research35 citesMar 25, 2019
- MarAdaCoSeg: Adaptive Shape Co-Segmentation with Group Consistency Lossno summary yetcs-cv1903.10297Adobe9 citesMar 25, 2019
- MarVideo Relationship Reasoning using Gated Spatio-Temporal Energy Graphno summary yetcs-cv1903.10547AllenAI14 citesMar 25, 2019
- MarLarge-scale interactive object segmentation with human annotatorsno summary yetcs-cv1903.10830Google Research17 citesMar 26, 2019
- MarGeneralized Feedback Loop for Joint Hand-Object Pose Estimationno summary yetcs-cv1903.10883Google Research2 citesMar 25, 2019
- MarBAE-NET: Branched Autoencoder for Shape Co-Segmentationno summary yetcs-cv1903.11228Adobe8 citesMar 27, 2019
- MarLaplace Landmark Localizationno summary yetcs-cv1903.11633Snap9 citesMar 27, 2019
- MarAlign2Ground: Weakly Supervised Phrase Grounding Guided by Image-Caption Alignmentno summary yetcs-cv1903.11649Meta / FAIR3 citesMar 27, 2019
- MarRefineLoc: Iterative Refinement for Weakly-Supervised Action Localizationno summary yetcs-cv1904.00227Adobe9 citesMar 30, 2019
- FebDifferentiable Grammars for Videosno summary yetcs-cv1902.00505Google Research0 citesFeb 1, 2019
- FebHierarchical Photo-Scene Encoder for Album Storytellingno summary yetcs-cv1902.00669Tencent2 citesFeb 2, 2019
- FebDeep-Emotion: Facial Expression Recognition Using Attentional Convolutional Networkno summary yetcs-cv1902.01019Snap127 citesFeb 4, 2019
- FebSkeleton-Based Online Action Prediction Using Scale Selection Networkno summary yetcs-cv1902.03084Alibaba2 citesFeb 8, 2019
- FebA 3D Probabilistic Deep Learning System for Detection and Diagnosis of Lung Cancer Using Low-Dose CT Scansno summary yetcs-cv1902.03233Google Research22 citesFeb 8, 2019
- FebAngle-Closure Detection in Anterior Segment OCT based on Multi-Level Deep Networkno summary yetcs-cv1902.03585Baidu80 citesFeb 10, 2019
- FebTaking a HINT: Leveraging Explanations to Make Vision and Language Models More Groundedno summary yetcs-cv1902.03751Meta / FAIR26 citesFeb 11, 2019
- FebMulti-Prototype Networks for Unconstrained Set-based Face Recognitionno summary yetcs-cv1902.04755Tencent6 citesFeb 13, 2019
- FebCycle-Consistency for Robust Visual Question Answeringno summary yetcs-cv1902.05660Meta / FAIR23 citesFeb 15, 2019
- FebTransfusion: Understanding Transfer Learning for Medical Imagingno summary yetcs-cv1902.07208Google Research643 citesFeb 14, 2019
- FebSpatio-Temporal Convolutional LSTMs for Tumor Growth Prediction by Learning 4D Longitudinal Patient Datano summary yetcs-cv1902.08716NVIDIA8 citesFeb 23, 2019
- FebBeyond Photometric Loss for Self-Supervised Ego-Motion Estimationno summary yetcs-cv1902.09103Tencent6 citesFeb 25, 2019
- FebDDFlow: Learning Optical Flow with Unlabeled Data Distillationno summary yetcs-cv1902.09145Tencent17 citesFeb 25, 2019
- FebDetecting Lesion Bounding Ellipses With Gaussian Proposal Networksno summary yetcs-cv1902.09658Baidu0 citesFeb 25, 2019
- FebHarmonic Unpaired Image-to-image Translationno summary yetcs-cv1902.09727Google Research30 citesFeb 26, 2019
- FebAn Annotation Saved is an Annotation Earned: Using Fully Synthetic Training for Object Instance Detectionno summary yetcs-cv1902.09967Google Research29 citesFeb 26, 2019
- FebGraph-RISE: Graph-Regularized Image Semantic Embeddingno summary yetcs-cv1902.10814Google Research29 citesFeb 14, 2019
- FebCircConv: A Structured Convolution with Low Complexityno summary yetcs-cv1902.11268Google Research1 citesFeb 28, 2019
- JanDetecting Text in the Wild with Deep Character Embedding Networkno summary yetcs-cv1901.00363Baidu2 citesJan 2, 2019
- JanAttribute-Aware Attention Model for Fine-grained Representation Learningno summary yetcs-cv1901.00392Alibaba2 citesJan 2, 2019
- JanLearning From Less Data: A Unified Data Subset Selection and Active Learning Framework for Computer Visionno summary yetcs-cv1901.01151Microsoft Research6 citesJan 3, 2019
- JanAdversarial Examples Versus Cloud-based Detectors: A Black-box Empirical Studyno summary yetcs-cv1901.01223Alibaba11 citesJan 4, 2019
- JanAVA-ActiveSpeaker: An Audio-Visual Dataset for Active Speaker Detectionno summary yetcs-cv1901.01342Google Research19 citesJan 5, 2019
- JanStereoscopic Dark Flash for Low-light Photographyno summary yetcs-cv1901.01370Google Research1 citesJan 5, 2019
- JanTencent ML-Images: A Large-Scale Multi-Label Image Database for Visual Representation Learningno summary yetcs-cv1901.01703Tencent79 citesJan 7, 2019
- JanHuman Pose Estimation with Spatial Contextual Informationno summary yetcs-cv1901.01760Baidu63 citesJan 7, 2019
- JanLow-Shot Learning from Imaginary 3D Modelno summary yetcs-cv1901.01868Amazon1 citesJan 4, 2019
- JanNormalized Object Coordinate Space for Category-Level 6D Object Pose and Size Estimationno summary yetcs-cv1901.02970Google Research44 citesJan 9, 2019
- JanThe Liver Tumor Segmentation Benchmark (LiTS)no summary yetcs-cv1901.04056IBM Research1,150 citesJan 13, 2019
- JaniPhys: An Open Non-Contact Imaging-Based Physiological Measurement Toolboxno summary yetcs-cv1901.04366Microsoft Research4 citesJan 14, 2019
- JanWhole-Slide Image Focus Quality: Automatic Assessment and Impact on AI Cancer Detectionno summary yetcs-cv1901.04619Google Research95 citesJan 15, 2019
- JanDomain Adaptation for Structured Output via Discriminative Patch Representationsno summary yetcs-cv1901.05427Google Research47 citesJan 16, 2019
- JanLearning single-image 3D reconstruction by generative modelling of shape, pose and shadingno summary yetcs-cv1901.06447Google Research10 citesJan 19, 2019
- JanLayoutGAN: Generating Graphic Layouts with Wireframe Discriminatorsno summary yetcs-cv1901.06767Adobe31 citesJan 21, 2019
- JanRead, Watch, and Move: Reinforcement Learning for Temporally Grounding Natural Language Descriptions in Videosno summary yetcs-cv1901.06829Baidu19 citesJan 21, 2019
- JanModeling Human Motion with Quaternion-based Neural Networksno summary yetcs-cv1901.07677Meta / FAIR141 citesJan 21, 2019
- JanAADS: Augmented Autonomous Driving Simulation using Data-driven Algorithmsno summary yetcs-cv1901.07849Baidu210 citesJan 23, 2019
- JanRevisiting Self-Supervised Visual Representation Learningno summary yetcs-cv1901.09005Google Research68 citesJan 25, 2019
- JanAudio-Visual Scene-Aware Dialogno summary yetcs-cv1901.09107Meta / FAIR35 citesJan 25, 2019
- JanSimilar Image Search for Histopathology: SMILYno summary yetcs-cv1901.11112Google Research126 citesJan 30, 2019
- JanMONet: Unsupervised Scene Decomposition and Representationno summary yetcs-cv1901.11390Google Research194 citesJan 22, 2019
- JanHotels-50K: A Global Hotel Recognition Datasetno summary yetcs-cv1901.11397Adobe0 citesJan 26, 2019
- JanSemantic Redundancies in Image-Classification Datasets: The 10% You Don't Needno summary yetcs-cv1901.11409Google Research18 citesJan 29, 2019
2018
258- DecExplaining the Ambiguity of Object Detection and 6D Pose From Visual Datano summary yetcs-cv1812.00287Google Research8 citesDec 1, 2018
- DecLearning to Learn How to Learn: Self-Adaptive Visual Navigation Using Meta-Learningno summary yetcs-cv1812.00971AllenAI16 citesDec 3, 2018
- DecTextField: Learning A Deep Direction Field for Irregular Scene Text Detectionno summary yetcs-cv1812.01393Alibaba365 citesDec 4, 2018
- DecMeta Learning Deep Visual Words for Fast Video Object Segmentationno summary yetcs-cv1812.01397Google Research9 citesDec 4, 2018
- DecThe Visual Centrifuge: Model-Free Layered Video Representationsno summary yetcs-cv1812.01461DeepMind1 citesDec 4, 2018
- DecGenerating High Fidelity Images with Subscale Pixel Networks and Multidimensional Upscalingno summary yetcs-cv1812.01608Google Research58 citesDec 4, 2018
- DecTowards Accurate Generative Models of Video: A New Metric & Challengesno summary yetcs-cv1812.01717Google Research194 citesDec 3, 2018
- DecInteractive Full Image Segmentation by Considering All Regions Jointlyno summary yetcs-cv1812.01888Google Research9 citesDec 5, 2018
- DecGuided Zoom: Questioning Network Evidence for Fine-grained Classificationno summary yetcs-cv1812.02626Adobe15 citesDec 6, 2018
- DecVideo Action Transformer Networkno summary yetcs-cv1812.02707DeepMind6 citesDec 6, 2018
- DecCross-Domain 3D Equivariant Image Embeddingsno summary yetcs-cv1812.02716Google Research8 citesDec 6, 2018
- DecSpatial Knowledge Distillation to aid Visual Reasoningno summary yetcs-cv1812.03631Adobe2 citesDec 10, 2018
- DecOccupancy Networks: Learning 3D Reconstruction in Function Spaceno summary yetcs-cv1812.03828Google Research116 citesDec 10, 2018
- DecWeakly Supervised Dense Event Captioning in Videosno summary yetcs-cv1812.03849Microsoft Research63 citesDec 10, 2018
- Dec2.5D Visual Soundno summary yetcs-cv1812.04204Meta / FAIR4 citesDec 11, 2018
- DecDomain-Aware SE Network for Sketch-based Image Retrieval with Multiplicative Euclidean Margin Softmaxno summary yetcs-cv1812.04275Baidu1 citesDec 11, 2018
- DecGrounded Human-Object Interaction Hotspots from Videono summary yetcs-cv1812.04558Meta / FAIR12 citesDec 11, 2018
- DecPyramid Network with Online Hard Example Mining for Accurate Left Atrium Segmentationno summary yetcs-cv1812.05802Tencent6 citesDec 14, 2018
- DecOn Attention Modules for Audio-Visual Synchronizationno summary yetcs-cv1812.06071Netflix4 citesDec 14, 2018
- DecEfficient Super Resolution Using Binarized Neural Networkno summary yetcs-cv1812.06378Adobe5 citesDec 16, 2018
- DecClassifier and Exemplar Synthesis for Zero-Shot Learningno summary yetcs-cv1812.06423Google Research5 citesDec 16, 2018
- DecUnsupervised Video Object Segmentation with Distractor-Aware Online Adaptationno summary yetcs-cv1812.07712Google Research1 citesDec 19, 2018
- DecD3D: Distilled 3D Networks for Video Action Recognitionno summary yetcs-cv1812.08249Google Research26 citesDec 19, 2018
- DecSequential Attention GAN for Interactive Image Editingno summary yetcs-cv1812.08352Microsoft Research7 citesDec 20, 2018
- DecThree Dimensional Reconstruction of Botanical Trees with Simulatable Geometryno summary yetcs-cv1812.08849NVIDIA0 citesDec 20, 2018
- DecDeep Learning and Glaucoma Specialists: The Relative Importance of Optic Disc Features to Predict Glaucoma Referral in Fundus Photosno summary yetcs-cv1812.08911Google Research185 citesDec 21, 2018
- DecSlimmable Neural Networksno summary yetcs-cv1812.08928Snap235 citesDec 21, 2018
- DecWireless Software Synchronization of Multiple Distributed Camerasno summary yetcs-cv1812.09366Google Research38 citesDec 21, 2018
- DecTextNet: Irregular Text Reading from Images with an End-to-End Trainable Networkno summary yetcs-cv1812.09900Baidu3 citesDec 24, 2018
- DecExploring the Challenges towards Lifelong Fact Learningno summary yetcs-cv1812.10524Meta / FAIR0 citesDec 26, 2018
- DecLarge-Scale Object Detection of Images from Network Cameras in Variable Ambient Lighting Conditionsno summary yetcs-cv1812.11901Meta / FAIR1 citesDec 31, 2018
- NovCariGANs: Unpaired Photo-to-Caricature Translationno summary yetcs-cv1811.00222Microsoft Research103 citesNov 1, 2018
- NovTowards Highly Accurate and Stable Face Alignment for High-Resolution Videosno summary yetcs-cv1811.00342Tencent3 citesNov 1, 2018
- NovLearning from Large-scale Noisy Web Data with Ubiquitous Reweighting for Image Classificationno summary yetcs-cv1811.00700Baidu3 citesNov 2, 2018
- NovThe Open Images Dataset V4: Unified image classification, object detection, and visual relationship detection at scaleno summary yetcs-cv1811.00982Google Research1,609 citesNov 2, 2018
- NovBi-Real Net: Binarizing Deep Network Towards Real-Network Performanceno summary yetcs-cv1811.01335Tencent8 citesNov 4, 2018
- NovStNet: Local and Global Spatial-Temporal Modeling for Action Recognitionno summary yetcs-cv1811.01549Baidu23 citesNov 5, 2018
- NovTrafficPredict: Trajectory Prediction for Heterogeneous Traffic-Agentsno summary yetcs-cv1811.02146Baidu72 citesNov 6, 2018
- NovSuper-Identity Convolutional Neural Network for Face Hallucinationno summary yetcs-cv1811.02328Tencent17 citesNov 6, 2018
- NovLookinGood: Enhancing Performance Capture with Real-time Neural Re-Renderingno summary yetcs-cv1811.05029Google Research16 citesNov 12, 2018
- NovDepth Prediction Without the Sensors: Leveraging Structure for Unsupervised Learning from Monocular Videosno summary yetcs-cv1811.06152Google Research58 citesNov 15, 2018
- NovDevelopment and Validation of a Deep Learning Algorithm for Improving Gleason Scoring of Prostate Cancerno summary yetcs-cv1811.06497Google Research498 citesNov 15, 2018
- NovAugmented LiDAR Simulator for Autonomous Drivingno summary yetcs-cv1811.07112Baidu155 citesNov 17, 2018
- NovRevisiting Image-Language Networks for Open-ended Phrase Detectionno summary yetcs-cv1811.07212NVIDIA17 citesNov 17, 2018
- NovCompressing Recurrent Neural Networks with Tensor Ring for Action Recognitionno summary yetcs-cv1811.07503Tencent13 citesNov 19, 2018
- NovScalable Logo Recognition using Proxiesno summary yetcs-cv1811.08009Amazon8 citesNov 19, 2018
- NovVisual Font Pairingno summary yetcs-cv1811.08015Adobe1 citesNov 19, 2018
- NovA Novel Integrated Framework for Learning both Text Detection and Recognitionno summary yetcs-cv1811.08611Alibaba1 citesNov 21, 2018
- NovJoint Face Hallucination and Deblurring via Structure Generation and Detail Enhancementno summary yetcs-cv1811.09019Tencent0 citesNov 22, 2018
- NovFast Object Class Labelling via Speechno summary yetcs-cv1811.09461Google Research0 citesNov 23, 2018
- NovTell, Draw, and Repeat: Generating and Modifying Images Based on Continual Linguistic Instructionno summary yetcs-cv1811.09845Microsoft Research11 citesNov 24, 2018
- NovA pooling based scene text proposal technique for scene text reading in the wildno summary yetcs-cv1811.10003Tencent22 citesNov 25, 2018
- NovEvolving Space-Time Neural Architectures for Videosno summary yetcs-cv1811.10636Google Research4 citesNov 26, 2018
- NovBilateral Adversarial Training: Towards Fast Training of More Robust Models Against Adversarial Attacksno summary yetcs-cv1811.10716Baidu8 citesNov 26, 2018
- NovProbabilistic Object Detection: Definition and Evaluationno summary yetcs-cv1811.10800Google Research16 citesNov 27, 2018
- NovFrom Recognition to Cognition: Visual Commonsense Reasoningno summary yetcs-cv1811.10830AllenAI59 citesNov 27, 2018
- NovUnprocessing Images for Learned Raw Denoisingno summary yetcs-cv1811.11127Google Research64 citesNov 27, 2018
- NovA Compact Embedding for Facial Expression Similarityno summary yetcs-cv1811.11283Google Research9 citesNov 27, 2018
- NovDeep Regionlets: Blended Representation and Deep Learning for Generic Object Detectionno summary yetcs-cv1811.11318Apple5 citesNov 28, 2018
- NovESPNetv2: A Light-weight, Power Efficient, and General Purpose Convolutional Neural Networkno summary yetcs-cv1811.11431AllenAI40 citesNov 28, 2018
- NovStrike (with) a Pose: Neural Networks Are Easily Fooled by Strange Poses of Familiar Objectsno summary yetcs-cv1811.11553Adobe45 citesNov 28, 2018
- Nov3D human pose estimation in video with temporal convolutions and semi-supervised trainingno summary yetcs-cv1811.11742Meta / FAIR4 citesNov 28, 2018
- NovLearning to Synthesize Motion Blurno summary yetcs-cv1811.11745Google Research0 citesNov 27, 2018
- NovJoint Correction of Attenuation and Scatter Using Deep Convolutional Neural Networks (DCNN) for Time-of-Flight PETno summary yetcs-cv1811.11852Microsoft Research70 citesNov 28, 2018
- NovDiscovering Spatio-Temporal Action Tubesno summary yetcs-cv1811.12248NVIDIA0 citesNov 29, 2018
- NovTouchdown: Natural Language Navigation and Spatial Reasoning in Visual Street Environmentsno summary yetcs-cv1811.12354Google Research46 citesNov 29, 2018
- NovClassification is a Strong Baseline for Deep Metric Learningno summary yetcs-cv1811.12649Amazon37 citesNov 30, 2018
- NovMicroscope 2.0: An Augmented Reality Microscope with Real-time Artificial Intelligence Integrationno summary yetcs-cv1812.00825Google Research342 citesNov 21, 2018
- OctLearning Depth with Convolutional Spatial Propagation Networkno summary yetcs-cv1810.02695Baidu37 citesOct 4, 2018
- OctContext-Aware Deep Spatio-Temporal Network for Hand Pose Estimation from Depth Imagesno summary yetcs-cv1810.02994Alibaba18 citesOct 6, 2018
- OctSanity Checks for Saliency Mapsno summary yetcs-cv1810.03292Google Research606 citesOct 8, 2018
- OctLocal Explanation Methods for Deep Neural Networks Lack Sensitivity to Parameter Valuesno summary yetcs-cv1810.03307Google Research22 citesOct 8, 2018
- OctOrthogonal Deep Features Decomposition for Age-Invariant Face Recognitionno summary yetcs-cv1810.07599Tencent10 citesOct 17, 2018
- OctDeepLens: Shallow Depth Of Field From A Single Imageno summary yetcs-cv1810.08100Adobe19 citesOct 18, 2018
- OctLearning Material-Aware Local Descriptors for 3D Shapesno summary yetcs-cv1810.08729Adobe13 citesOct 20, 2018
- OctImage Inpainting via Generative Multi-column Convolutional Neural Networksno summary yetcs-cv1810.08771Tencent193 citesOct 20, 2018
- OctPredicting optical coherence tomography-derived diabetic macular edema grades from fundus photographs using deep learningno summary yetcs-cv1810.10342DeepMind134 citesOct 18, 2018
- OctNeighbourhood Consensus Networksno summary yetcs-cv1810.10510DeepMind153 citesOct 24, 2018
- OctUnsupervised Multi-Target Domain Adaptation: An Information Theoretic Approachno summary yetcs-cv1810.11547DeepMind27 citesOct 26, 2018
- Oct3D MRI brain tumor segmentation using autoencoder regularizationno summary yetcs-cv1810.11654NVIDIA61 citesOct 27, 2018
- OctRecurrent Transformer Networks for Semantic Correspondenceno summary yetcs-cv1810.12155Microsoft Research50 citesOct 29, 2018
- OctReal-Time RGB-D Camera Pose Estimation in Novel Scenes using a Relocalisation Cascadeno summary yetcs-cv1810.12163Google Research86 citesOct 29, 2018
- OctPyramidal Person Re-IDentification via Multi-Loss Dynamic Trainingno summary yetcs-cv1810.12193Tencent15 citesOct 29, 2018
- OctDeepGRU: Deep Gesture Recognition Utilityno summary yetcs-cv1810.12514NVIDIA2 citesOct 30, 2018
- OctDropBlock: A regularization method for convolutional networksno summary yetcs-cv1810.12890Google Research518 citesOct 30, 2018
- OctCompact Generalized Non-local Networkno summary yetcs-cv1810.13125Baidu68 citesOct 31, 2018
- OctComputational Histological Staining and Destaining of Prostate Core Biopsy RGB Images with Generative Adversarial Neural Networksno summary yetcs-cv1811.02642IBM Research60 citesOct 26, 2018
- SepYouTube-VOS: Sequence-to-Sequence Video Object Segmentationno summary yetcs-cv1809.00461Snap20 citesSep 3, 2018
- SepLocalizing Moments in Video with Temporal Languageno summary yetcs-cv1809.01337Adobe143 citesSep 5, 2018
- SepSemantic Human Mattingno summary yetcs-cv1809.01354Alibaba0 citesSep 5, 2018
- SepVisual Coreference Resolution in Visual Dialog using Neural Module Networksno summary yetcs-cv1809.01816Meta / FAIR0 citesSep 6, 2018
- SepSurface Light Field Fusionno summary yetcs-cv1809.02057Adobe0 citesSep 6, 2018
- SepDeep Audio-Visual Speech Recognitionno summary yetcs-cv1809.02108DeepMind781 citesSep 6, 2018
- SepRepresenting Images in 200 Bytes: Compression via Triangulationno summary yetcs-cv1809.02257Google Research1 citesSep 7, 2018
- SepJoint Autoregressive and Hierarchical Priors for Learned Image Compressionno summary yetcs-cv1809.02736Google Research594 citesSep 8, 2018
- SepSearching for Efficient Multi-Scale Architectures for Dense Image Predictionno summary yetcs-cv1809.04184Google Research40 citesSep 11, 2018
- SepAdaptive Sampling Towards Fast Graph Representation Learningno summary yetcs-cv1809.05343Tencent230 citesSep 14, 2018
- SepA Framework towards Domain Specific Video Summarizationno summary yetcs-cv1809.08854Microsoft Research3 citesSep 24, 2018
- SepDeformable Object Tracking with Gated Fusionno summary yetcs-cv1809.10417Tencent36 citesSep 27, 2018
- SepInterest point detectors stability evaluation on ApolloScape datasetno summary yetcs-cv1809.11039Google Research1 citesSep 28, 2018
- SepSConE: Siamese Constellation Embedding Descriptor for Image Matchingno summary yetcs-cv1809.11054Google Research0 citesSep 28, 2018
- SepNon-local NetVLAD Encoding for Video Classificationno summary yetcs-cv1810.00207Tencent9 citesSep 29, 2018
- AugDepth Estimation via Affinity Learned with Convolutional Spatial Propagation Networkno summary yetcs-cv1808.00150Baidu1 citesAug 1, 2018
- AugLearning Visual Question Answering by Bootstrapping Hard Attentionno summary yetcs-cv1808.00300DeepMind9 citesAug 1, 2018
- AugPhysics-Based Generative Adversarial Models for Image Restoration and Beyondno summary yetcs-cv1808.00605Tencent32 citesAug 2, 2018
- AugLearning Actionable Representations from Visual Observationsno summary yetcs-cv1808.00928Google Research16 citesAug 2, 2018
- AugComposition Loss for Counting, Density Map Estimation and Localization in Dense Crowdsno summary yetcs-cv1808.01050NVIDIA61 citesAug 2, 2018
- AugParsing Geometry Using Structure-Aware Shape Templatesno summary yetcs-cv1808.01337Google Research4 citesAug 3, 2018
- AugRethinking Pose in 3D: Multi-stage Refinement and Recovery for Markerless Motion Captureno summary yetcs-cv1808.01525Amazon3 citesAug 4, 2018
- AugVideo Re-localizationno summary yetcs-cv1808.01575Tencent3 citesAug 5, 2018
- AugContemplating Visual Emotions: Understanding and Overcoming Dataset Biasno summary yetcs-cv1808.02212Adobe3 citesAug 7, 2018
- AugChoose Your Neuron: Incorporating Domain Knowledge through Neuron-Importanceno summary yetcs-cv1808.02861Meta / FAIR5 citesAug 8, 2018
- AugEnd-to-end Active Object Tracking and Its Real-world Deployment via Reinforcement Learningno summary yetcs-cv1808.03405Tencent6 citesAug 10, 2018
- AugFacial Action Unit Detection Using Attention and Relation Learningno summary yetcs-cv1808.03457Tencent89 citesAug 10, 2018
- AugFully-Automated Analysis of Body Composition from CT in Cancer Patients Using Convolutional Neural Networksno summary yetcs-cv1808.03844IBM Research43 citesAug 11, 2018
- AugVisual Reasoning with Multi-hop Feature Modulationno summary yetcs-cv1808.04446Google Research25 citesAug 3, 2018
- AugRecycle-GAN: Unsupervised Video Retargetingno summary yetcs-cv1808.05174Meta / FAIR15 citesAug 15, 2018
- AugAnatomyNet: Deep Learning for Fast and Fully Automated Whole-volume Segmentation of Head and Neck Anatomyno summary yetcs-cv1808.05238Tencent526 citesAug 15, 2018
- AugMedical Image Imputation from Image Collectionsno summary yetcs-cv1808.05732Google Research41 citesAug 17, 2018
- AugConcept Mask: Large-Scale Segmentation from Semantic Conceptsno summary yetcs-cv1808.06032Meta / FAIR3 citesAug 18, 2018
- AugDynamic Temporal Alignment of Speech to Lipsno summary yetcs-cv1808.06250Google Research1 citesAug 19, 2018
- AugVideo-to-Video Synthesisno summary yetcs-cv1808.06601Adobe127 citesAug 20, 2018
- AugConstrained-size Tensorflow Models for YouTube-8M Video Understanding Challengeno summary yetcs-cv1808.06739Google Research0 citesAug 21, 2018
- AugCan 3D Pose be Learned from 2D Projections Alone?no summary yetcs-cv1808.07182Amazon3 citesAug 22, 2018
- AugLearning Hierarchical Semantic Image Manipulation through Structured Representationsno summary yetcs-cv1808.07535Google Research12 citesAug 22, 2018
- AugARBEE: Towards Automated Recognition of Bodily Expression of Emotion In the Wildno summary yetcs-cv1808.09568Amazon97 citesAug 28, 2018
- AugTask adapted reconstruction for inverse problemsno summary yetcs-cv1809.00948DeepMind15 citesAug 27, 2018
- JulVolumetric performance capture from minimal camera viewpointsno summary yetcs-cv1807.01950Adobe1 citesJul 5, 2018
- JulDetecting Visual Relationships Using Box Attentionno summary yetcs-cv1807.02136Google Research18 citesJul 5, 2018
- JulRepresenting a Partially Observed Non-Rigid 3D Human Using Eigen-Texture and Eigen-Deformationno summary yetcs-cv1807.02632Microsoft Research0 citesJul 7, 2018
- JulDiscovery of Latent 3D Keypoints via End-to-end Geometric Reasoningno summary yetcs-cv1807.03146Google Research132 citesJul 5, 2018
- JulVideo Captioning with Boundary-aware Hierarchical Language Decoding and Joint Video Predictionno summary yetcs-cv1807.03658Adobe3 citesJul 8, 2018
- JulEffective Use of Synthetic Data for Urban Scene Semantic Segmentationno summary yetcs-cv1807.06132NVIDIA19 citesJul 16, 2018
- JulBAM: Bottleneck Attention Moduleno summary yetcs-cv1807.06514Adobe48 citesJul 17, 2018
- JulCBAM: Convolutional Block Attention Moduleno summary yetcs-cv1807.06521Adobe336 citesJul 17, 2018
- JulBridging the Accuracy Gap for 2-bit Quantized Neural Networks (QNN)no summary yetcs-cv1807.06964Google Research38 citesJul 17, 2018
- JulHybrid Scene Compression for Visual Localizationno summary yetcs-cv1807.07512Microsoft Research2 citesJul 19, 2018
- JulPerson Search via A Mask-Guided Two-Stream CNN Modelno summary yetcs-cv1807.08107Tencent24 citesJul 21, 2018
- JulTowards Privacy-Preserving Visual Recognition via Adversarial Training: A Pilot Studyno summary yetcs-cv1807.08379Adobe18 citesJul 22, 2018
- JulStereoNet: Guided Hierarchical Refinement for Real-Time Edge-Aware Depth Predictionno summary yetcs-cv1807.08865Google Research13 citesJul 24, 2018
- JulVisual Dynamics: Stochastic Future Generation via Layered Cross Convolutional Networksno summary yetcs-cv1807.09245Google Research34 citesJul 24, 2018
- JulRecurrent Fusion Network for Image Captioningno summary yetcs-cv1807.09986Tencent1 citesJul 26, 2018
- JulLQ-Nets: Learned Quantization for Highly Accurate and Compact Deep Neural Networksno summary yetcs-cv1807.10029Microsoft Research45 citesJul 26, 2018
- JulSuperpixel Sampling Networksno summary yetcs-cv1807.10174NVIDIA0 citesJul 26, 2018
- JulLayer-structured 3D Scene Inference via View Synthesisno summary yetcs-cv1807.10264Google Research5 citesJul 26, 2018
- JulPairwise Body-Part Attention for Recognizing Human-Object Interactionsno summary yetcs-cv1807.10889Tencent8 citesJul 28, 2018
- JulSidekick Policy Learning for Active Visual Explorationno summary yetcs-cv1807.11010Meta / FAIR0 citesJul 29, 2018
- JulMnasNet: Platform-Aware Neural Architecture Search for Mobileno summary yetcs-cv1807.11626Google Research222 citesJul 31, 2018
- JunIGCV3: Interleaved Low-Rank Group Convolutions for Efficient Deep Neural Networksno summary yetcs-cv1806.00178Microsoft Research27 citesJun 1, 2018
- JunCFCM: Segmentation via Coarse to Fine Context Memoryno summary yetcs-cv1806.01413NVIDIA4 citesJun 4, 2018
- JunFocal Visual-Text Attention for Visual Question Answeringno summary yetcs-cv1806.01873Meta / FAIR76 citesJun 5, 2018
- JunFree-Form Image Inpainting with Gated Convolutionno summary yetcs-cv1806.03589Adobe99 citesJun 10, 2018
- JunMassively Parallel Video Networksno summary yetcs-cv1806.03863DeepMind2 citesJun 11, 2018
- JunSynthetic Depth-of-Field with a Single-Camera Mobile Phoneno summary yetcs-cv1806.04171Google Research184 citesJun 11, 2018
- JunLearning Visual Knowledge Memory Networks for Visual Question Answeringno summary yetcs-cv1806.04860Tencent14 citesJun 13, 2018
- JunA Probabilistic U-Net for Segmentation of Ambiguous Imagesno summary yetcs-cv1806.05034Google Research151 citesJun 13, 2018
- Jun3D-CODED : 3D Correspondences by Deep Deformationno summary yetcs-cv1806.05228Adobe293 citesJun 13, 2018
- JunFinding your Lookalike: Measuring Face Similarity Rather than Face Identityno summary yetcs-cv1806.05252Google Research0 citesJun 13, 2018
- JunReConvNet: Video Object Segmentation with Spatio-Temporal Features Modulationno summary yetcs-cv1806.05510Google Research0 citesJun 14, 2018
- JunUnsupervised Training for 3D Morphable Model Regressionno summary yetcs-cv1806.06098Google Research18 citesJun 15, 2018
- JunError Compensated Quantized SGD and its Applications to Large-scale Distributed Optimizationno summary yetcs-cv1806.08054Tencent38 citesJun 21, 2018
- JunDPP-Net: Device-aware Progressive Search for Pareto-optimal Neural Architecturesno summary yetcs-cv1806.08198Google Research18 citesJun 21, 2018
- JunSemi-Automatic RECIST Labeling on CT Scans with Cascaded Convolutional Neural Networksno summary yetcs-cv1806.09507NVIDIA4 citesJun 25, 2018
- JunTracking Emerges by Colorizing Videosno summary yetcs-cv1806.09594Google Research22 citesJun 25, 2018
- JunLearn-to-Score: Efficient 3D Scene Exploration by Predicting View Utilityno summary yetcs-cv1806.10354Microsoft Research1 citesJun 27, 2018
- MayExploring the Limits of Weakly Supervised Pretrainingno summary yetcs-cv1805.00932Meta / FAIR204 citesMay 2, 2018
- MayDeep Ordinal Hashing with Spatial Attentionno summary yetcs-cv1805.02459Meta / FAIR94 citesMay 7, 2018
- MayImage Retrieval with Mixed Initiative and Multimodal Feedbackno summary yetcs-cv1805.03134Snap5 citesMay 8, 2018
- MayWeakly and Semi Supervised Human Body Part Parsing via Pose-Guided Knowledge Transferno summary yetcs-cv1805.04310Tencent7 citesMay 11, 2018
- MayRevisiting Dilated Convolution: A Simple Approach for Weakly- and Semi- Supervised Semantic Segmentationno summary yetcs-cv1805.04574Tencent45 citesMay 11, 2018
- MayLearning Rich Features for Image Manipulation Detectionno summary yetcs-cv1805.04953Adobe82 citesMay 13, 2018
- MayOn Learning Associations of Faces and Voicesno summary yetcs-cv1805.05553Adobe21 citesMay 15, 2018
- MayVisual Representations for Semantic Target Driven Navigationno summary yetcs-cv1805.06066Google Research36 citesMay 15, 2018
- MayRobust and Efficient Graph Correspondence Transfer for Person Re-identificationno summary yetcs-cv1805.06323Tencent1 citesMay 15, 2018
- MayImproving Image Captioning with Conditional Generative Adversarial Netsno summary yetcs-cv1805.07112Tencent77 citesMay 18, 2018
- MayDeepPhys: Video-Based Physiological Measurement Using Convolutional Attention Networksno summary yetcs-cv1805.07888Microsoft Research44 citesMay 21, 2018
- MaySelf-supervised Multi-view Person Association and Its Applicationsno summary yetcs-cv1805.08717Amazon26 citesMay 22, 2018
- MayEfficient Relaxations for Dense CRFs with Sparse Higher Order Potentialsno summary yetcs-cv1805.09028Google Research2 citesMay 23, 2018
- MayExcitation Dropout: Encouraging Plasticity in Deep Neural Networksno summary yetcs-cv1805.09092Adobe7 citesMay 23, 2018
- MaySOSELETO: A Unified Approach to Transfer Learning and Training with Noisy Labelsno summary yetcs-cv1805.09622Google Research7 citesMay 24, 2018
- MayStereo Magnification: Learning View Synthesis using Multiplane Imagesno summary yetcs-cv1805.09817Google Research140 citesMay 24, 2018
- MayLearning from Multi-domain Artistic Images for Arbitrary Style Transferno summary yetcs-cv1805.09987Adobe8 citesMay 25, 2018
- MayPyramid Attention Network for Semantic Segmentationno summary yetcs-cv1805.10180Tencent235 citesMay 25, 2018
- MayUsing Syntax to Ground Referring Expressions in Natural Imagesno summary yetcs-cv1805.10547Microsoft Research35 citesMay 26, 2018
- MayImage-Dependent Local Entropy Models for Learned Image Compressionno summary yetcs-cv1805.12295Google Research3 citesMay 31, 2018
- MayDeepMiner: Discovering Interpretable Representations for Mammogram Classification and Explanationno summary yetcs-cv1805.12323Microsoft Research19 citesMay 31, 2018
- AprEnd-to-End Learning of Motion Representation for Video Understandingno summary yetcs-cv1804.00413Tencent31 citesApr 2, 2018
- Apr3D Interpreter Networks for Viewer-Centered Wireframe Modelingno summary yetcs-cv1804.00782Meta / FAIR19 citesApr 3, 2018
- AprLeft-Right Comparative Recurrent Model for Stereo Matchingno summary yetcs-cv1804.00796Tencent12 citesApr 3, 2018
- AprEnd-to-End Dense Video Captioning with Masked Transformerno summary yetcs-cv1804.00819Salesforce42 citesApr 3, 2018
- AprFine-grained Video Attractiveness Prediction Using Multimodal Deep Learning on a Large Real-world Datasetno summary yetcs-cv1804.01373Tencent11 citesApr 4, 2018
- AprPixel2Mesh: Generating 3D Mesh Models from Single RGB Imagesno summary yetcs-cv1804.01654Tencent140 citesApr 5, 2018
- AprLearning to Separate Object Sounds by Watching Unlabeled Videono summary yetcs-cv1804.01665Meta / FAIR28 citesApr 5, 2018
- AprLook into Person: Joint Body Parsing & Pose Estimation Network and A New Benchmarkno summary yetcs-cv1804.01984Adobe11 citesApr 5, 2018
- AprNoise-resistant Deep Learning for Object Classification in 3D Point Clouds Using a Point Pair Descriptorno summary yetcs-cv1804.02077Baidu26 citesApr 5, 2018
- AprLearning-based Video Motion Magnificationno summary yetcs-cv1804.02684Google Research13 citesApr 8, 2018
- AprGenerative Adversarial Networks for Extreme Learned Image Compressionno summary yetcs-cv1804.02958Google Research70 citesApr 9, 2018
- AprImagine This! Scripts to Compositions to Videosno summary yetcs-cv1804.03608AllenAI3 citesApr 10, 2018
- AprDecoupled Novel Object Captionerno summary yetcs-cv1804.03803Google Research70 citesApr 11, 2018
- AprDistort-and-Recover: Color Enhancement using Deep Reinforcement Learningno summary yetcs-cv1804.04450Adobe9 citesApr 12, 2018
- AprMultimodal Unsupervised Image-to-Image Translationno summary yetcs-cv1804.04732NVIDIA298 citesApr 12, 2018
- AprBodyNet: Volumetric Inference of 3D Human Body Shapesno summary yetcs-cv1804.04875Adobe426 citesApr 13, 2018
- AprMulti-level Semantic Feature Augmentation for One-shot Learningno summary yetcs-cv1804.05298Google Research247 citesApr 15, 2018
- AprNeural Kinematic Networks for Unsupervised Motion Retargettingno summary yetcs-cv1804.05653Adobe26 citesApr 16, 2018
- AprDCAN: Dual Channel-wise Alignment Networks for Unsupervised Scene Adaptationno summary yetcs-cv1804.05827Meta / FAIR22 citesApr 16, 2018
- AprA Fusion Framework for Camouflaged Moving Foreground Detection in the Wavelet Domainno summary yetcs-cv1804.05984Microsoft Research66 citesApr 16, 2018
- AprPlaneNet: Piece-wise Planar Reconstruction from a Single RGB Imageno summary yetcs-cv1804.06278Adobe15 citesApr 17, 2018
- AprImage Inpainting for Irregular Holes Using Partial Convolutionsno summary yetcs-cv1804.07723NVIDIA18 citesApr 20, 2018
- AprLarge Scale Scene Text Verification with Guided Attentionno summary yetcs-cv1804.08588Google Research0 citesApr 23, 2018
- AprSwitchable Temporal Propagation Networkno summary yetcs-cv1804.08758NVIDIA4 citesApr 23, 2018
- AprActor and Observer: Joint Modeling of First and Third-Person Videosno summary yetcs-cv1804.09627AllenAI141 citesApr 25, 2018
- MarPath Aggregation Network for Instance Segmentationno summary yetcs-cv1803.01534Tencent350 citesMar 5, 2018
- MarImproving the Improved Training of Wasserstein GANs: A Consistency Term and Its Dual Effectno summary yetcs-cv1803.01541Tencent52 citesMar 5, 2018
- MarST-GAN: Spatial Transformer Generative Adversarial Networks for Image Compositingno summary yetcs-cv1803.01837Adobe39 citesMar 5, 2018
- MarPersonalized Exposure Control Using Adaptive Metering and Reinforcement Learningno summary yetcs-cv1803.02269Microsoft Research0 citesMar 6, 2018
- MarLocal Kernels that Approximate Bayesian Regularization and Proximal Operatorsno summary yetcs-cv1803.03711Google Research8 citesMar 9, 2018
- MarKnowledge Aided Consistency for Weakly Supervised Phrase Groundingno summary yetcs-cv1803.03879Adobe11 citesMar 11, 2018
- MarExpert identification of visual primitives used by CNNs during mammogram classificationno summary yetcs-cv1803.04858Microsoft Research2 citesMar 13, 2018
- MarLEGO: Learning Edge with Geometry all at Once by Watching Videosno summary yetcs-cv1803.05648Baidu21 citesMar 15, 2018
- MarThe ApolloScape Open Dataset for Autonomous Driving and its Applicationno summary yetcs-cv1803.06184Baidu601 citesMar 16, 2018
- MarLearning to Segment via Cut-and-Pasteno summary yetcs-cv1803.06414Google Research5 citesMar 16, 2018
- MarReal-time Burst Photo Selection Using a Light-Head Adversarial Networkno summary yetcs-cv1803.07212Microsoft Research2 citesMar 20, 2018
- MarHierarchical Metric Learning and Matching for 2D and 3D Geometric Correspondencesno summary yetcs-cv1803.07231Google Research2 citesMar 20, 2018
- MarEnd-to-End Video Captioning with Multitask Reinforcement Learningno summary yetcs-cv1803.07950Tencent6 citesMar 21, 2018
- MarStacked Cross Attention for Image-Text Matchingno summary yetcs-cv1803.08024Microsoft Research48 citesMar 21, 2018
- MarPersonLab: Person Pose Estimation and Instance Segmentation with a Bottom-Up, Part-Based, Geometric Embedding Modelno summary yetcs-cv1803.08225Google Research35 citesMar 22, 2018
- MarOn Regularized Losses for Weakly-supervised CNN Segmentationno summary yetcs-cv1803.09569Adobe20 citesMar 26, 2018
- MarSingle Day Outdoor Photometric Stereono summary yetcs-cv1803.10850Adobe20 citesMar 28, 2018
- MarGenerative Modeling using the Sliced Wasserstein Distanceno summary yetcs-cv1803.11188Snap28 citesMar 29, 2018
- MarReconstruction Network for Video Captioningno summary yetcs-cv1803.11438Tencent41 citesMar 30, 2018
- MarTagging like Humans: Diverse and Distinct Image Annotationno summary yetcs-cv1804.00113Tencent8 citesMar 31, 2018
- MarMulti-label Learning with Missing Labels using Mixed Dependency Graphsno summary yetcs-cv1804.00117Tencent5 citesMar 31, 2018
- MarSnap Angle Prediction for 360$^{\circ}$ Panoramasno summary yetcs-cv1804.00126Meta / FAIR0 citesMar 31, 2018
- MarDeepIM: Deep Iterative Matching for 6D Pose Estimationno summary yetcs-cv1804.00175NVIDIA203 citesMar 31, 2018
- MarAdversarial Spatio-Temporal Learning for Video Deblurringno summary yetcs-cv1804.00533Tencent296 citesMar 28, 2018
- FebExplaining First Impressions: Modeling, Recognizing, and Explaining Apparent Personality from Videosno summary yetcs-cv1802.00745Microsoft Research34 citesFeb 2, 2018
- FebEfficient Video Object Segmentation via Network Modulationno summary yetcs-cv1802.01218Snap40 citesFeb 4, 2018
- FebEncoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentationno summary yetcs-cv1802.02611Google Research4,630 citesFeb 7, 2018
- FebGoing Deeper in Spiking Neural Networks: VGG and Residual Architecturesno summary yetcs-cv1802.02627Meta / FAIR59 citesFeb 7, 2018
- FebSpatially adaptive image compression using a tiled deep networkno summary yetcs-cv1802.02629Google Research6 citesFeb 7, 2018
- FebTexture Segmentation Based Video Compression Using Convolutional Neural Networksno summary yetcs-cv1802.02992Google Research0 citesFeb 8, 2018
- FebAMC: AutoML for Model Compression and Acceleration on Mobile Devicesno summary yetcs-cv1802.03494Google Research1,267 citesFeb 10, 2018
- FebWeb-Scale Responsive Visual Search at Bingno summary yetcs-cv1802.04914Microsoft Research10 citesFeb 14, 2018
- FebAtlasNet: A Papier-Mâché Approach to Learning 3D Surface Generationno summary yetcs-cv1802.05384Adobe178 citesFeb 15, 2018
- FebUnsupervised Learning of Depth and Ego-Motion from Monocular Video Using 3D Geometric Constraintsno summary yetcs-cv1802.05522Google Research1 citesFeb 15, 2018
- FebFast, Trainable, Multiscale Denoisingno summary yetcs-cv1802.06130Google Research0 citesFeb 16, 2018
- FebGlobal Pose Estimation with an Attention-based Recurrent Networkno summary yetcs-cv1802.06857Apple16 citesFeb 19, 2018
- FebChatPainter: Improving Text to Image Generation using Dialogueno summary yetcs-cv1802.08216Microsoft Research17 citesFeb 22, 2018
- FebConstrained Image Generation Using Binarized Neural Networks with Decision Proceduresno summary yetcs-cv1802.08795Microsoft Research0 citesFeb 24, 2018
- FebDetecting Comma-shaped Clouds for Severe Weather Forecasting using Shape and Motionno summary yetcs-cv1802.08937Amazon18 citesFeb 25, 2018
- JanInstance Embedding Transfer to Unsupervised Video Object Segmentationno summary yetcs-cv1801.00908Google Research14 citesJan 3, 2018
- JanMobileNetV2: Inverted Residuals and Linear Bottlenecksno summary yetcs-cv1801.04381Google Research2,274 citesJan 13, 2018
- JanFrame-Recurrent Video Super-Resolutionno summary yetcs-cv1801.04590Google Research50 citesJan 14, 2018
- JanSemi-supervised FusedGAN for Conditional Image Generationno summary yetcs-cv1801.05551Microsoft Research10 citesJan 17, 2018
- JanNDDR-CNN: Layerwise Feature Fusing in Multi-Task CNNs by Neural Discriminative Dimensionality Reductionno summary yetcs-cv1801.08297Tencent20 citesJan 25, 2018
- JanEfficient Hierarchical Graph-Based Segmentation of RGBD Videosno summary yetcs-cv1801.08981Microsoft Research52 citesJan 26, 2018
- JanDeepLung: Deep 3D Dual Path Nets for Automated Pulmonary Nodule Detection and Classificationno summary yetcs-cv1801.09555Baidu64 citesJan 25, 2018
- JanAction Recognition with Spatio-Temporal Visual Attention on Skeleton Image Sequencesno summary yetcs-cv1801.10304Snap2 citesJan 31, 2018
2017
167- DecDeformable Shape Completion with Graph Convolutional Autoencodersno summary yetcs-cv1712.00268Google Research23 citesDec 1, 2017
- DecDon't Just Assume; Look and Answer: Overcoming Priors for Visual Question Answeringno summary yetcs-cv1712.00377AllenAI40 citesDec 1, 2017
- DecMulti-Content GAN for Few-Shot Font Style Transferno summary yetcs-cv1712.00516Adobe30 citesDec 1, 2017
- DecProgressive Neural Architecture Searchno summary yetcs-cv1712.00559Google Research166 citesDec 2, 2017
- DecSemi-Global Stereo Matching with Surface Orientation Priorsno summary yetcs-cv1712.00818Microsoft Research7 citesDec 3, 2017
- DecA Perceptual Measure for Deep Single Image Camera Calibrationno summary yetcs-cv1712.01259Adobe12 citesDec 2, 2017
- DecSelf-supervised Learning of Motion Captureno summary yetcs-cv1712.01337Adobe133 citesDec 4, 2017
- DecAdversarial Attribute-Image Person Re-identificationno summary yetcs-cv1712.01493Tencent6 citesDec 5, 2017
- DecStructured Set Matching Networks for One-Shot Part Labelingno summary yetcs-cv1712.01867AllenAI3 citesDec 5, 2017
- DecSeparating Reflection and Transmission Images in the Wildno summary yetcs-cv1712.02099NVIDIA7 citesDec 6, 2017
- DecLearned Perceptual Image Enhancementno summary yetcs-cv1712.02864Google Research2 citesDec 7, 2017
- DecIQA: Visual Question Answering in Interactive Environmentsno summary yetcs-cv1712.03316AllenAI37 citesDec 9, 2017
- DecEye In-Painting with Exemplar Generative Adversarial Networksno summary yetcs-cv1712.03999Meta / FAIR12 citesDec 11, 2017
- DecCharacter-Based Handwritten Text Transcription with Attention Networksno summary yetcs-cv1712.04046NVIDIA30 citesDec 11, 2017
- DecLearning a Complete Image Indexing Pipelineno summary yetcs-cv1712.04480Amazon14 citesDec 12, 2017
- DecRethinking Spatiotemporal Feature Learning: Speed-Accuracy Trade-offs in Video Classificationno summary yetcs-cv1712.04851Google Research3 citesDec 13, 2017
- DecUnsupervised Domain Adaptation for 3D Keypoint Estimation via View Consistencyno summary yetcs-cv1712.05765Snap0 citesDec 15, 2017
- DecDeep Burst Denoisingno summary yetcs-cv1712.05790Meta / FAIR5 citesDec 15, 2017
- DecObjects that Soundno summary yetcs-cv1712.06651DeepMind18 citesDec 18, 2017
- DecEnd-to-end weakly-supervised semantic alignmentno summary yetcs-cv1712.06861DeepMind6 citesDec 19, 2017
- DecDeep learning for predicting refractive error from retinal fundus imagesno summary yetcs-cv1712.07798DeepMind183 citesDec 21, 2017
- DecOn the Integration of Optical Flow and Action Recognitionno summary yetcs-cv1712.08416Meta / FAIR61 citesDec 22, 2017
- NovMitigating Adversarial Effects Through Randomizationno summary yetcs-cv1711.01991Baidu196 citesNov 6, 2017
- NovSaliency Prediction for Mobile User Interfacesno summary yetcs-cv1711.03726Adobe0 citesNov 10, 2017
- NovGoing Further with Point Pair Featuresno summary yetcs-cv1711.04061Google Research127 citesNov 11, 2017
- NovLatent Constrained Correlation Filterno summary yetcs-cv1711.04192Microsoft Research49 citesNov 11, 2017
- NovCrowd counting via scale-adaptive convolutional neural networkno summary yetcs-cv1711.04433Tencent23 citesNov 13, 2017
- NovXGAN: Unsupervised Image-to-Image Translation for Many-to-Many Mappingsno summary yetcs-cv1711.05139DeepMind68 citesNov 14, 2017
- NovDynamic Zoom-in Network for Fast Object Detection in Large Imagesno summary yetcs-cv1711.05187DeepMind15 citesNov 14, 2017
- NovC-WSL: Count-guided Weakly Supervised Localizationno summary yetcs-cv1711.05282DeepMind9 citesNov 14, 2017
- NovRobust and Precise Vehicle Localization based on Multi-sensor Fusion in Diverse City Scenesno summary yetcs-cv1711.05805Baidu287 citesNov 15, 2017
- NovNISP: Pruning Networks using Neuron Importance Score Propagationno summary yetcs-cv1711.05908DeepMind62 citesNov 16, 2017
- NovLanguage-Based Image Editing with Recurrent Attentive Modelsno summary yetcs-cv1711.06288Microsoft Research9 citesNov 16, 2017
- NovLook, Imagine and Match: Improving Textual-Visual Cross-Modal Retrieval with Generative Modelsno summary yetcs-cv1711.06420Alibaba37 citesNov 17, 2017
- NovExcitation Backprop for RNNsno summary yetcs-cv1711.06778Adobe0 citesNov 18, 2017
- NovAperture Supervision for Monocular Depth Estimationno summary yetcs-cv1711.07933Google Research3 citesNov 21, 2017
- NovDeep Video Generation, Prediction and Completion of Human Action Sequencesno summary yetcs-cv1711.08682Tencent83 citesNov 23, 2017
- NovReal-Time Seamless Single Shot 6D Object Pose Predictionno summary yetcs-cv1711.08848Microsoft Research15 citesNov 24, 2017
- NovImage Generation from Sketch Constraint Using Contextual GANno summary yetcs-cv1711.08972Tencent2 citesNov 24, 2017
- NovVideo Enhancement with Task-Oriented Flowno summary yetcs-cv1711.09078Google Research1,303 citesNov 24, 2017
- NovAttention Clusters: Purely Attention Based Local Feature Integration for Video Classificationno summary yetcs-cv1711.09550Baidu19 citesNov 27, 2017
- NovHP-GAN: Probabilistic 3D human motion prediction via GANno summary yetcs-cv1711.09561Microsoft Research24 citesNov 27, 2017
- NovMemory Aware Synapses: Learning what (not) to forgetno summary yetcs-cv1711.09601Meta / FAIR43 citesNov 27, 2017
- NovRecurrent Segmentation for Variable Computational Budgetsno summary yetcs-cv1711.10151Google Research6 citesNov 28, 2017
- NovBLADE: Filter Learning for General Purpose Computational Photographyno summary yetcs-cv1711.10700Google Research0 citesNov 29, 2017
- NovAutomatic Generation of Constrained Furniture Layoutsno summary yetcs-cv1711.10939Google Research14 citesNov 29, 2017
- NovConvolutional Networks with Adaptive Inference Graphsno summary yetcs-cv1711.11503Google Research13 citesNov 30, 2017
- NovHigh-Resolution Image Synthesis and Semantic Manipulation with Conditional GANsno summary yetcs-cv1711.11585NVIDIA301 citesNov 30, 2017
- NovToward Multimodal Image-to-Image Translationno summary yetcs-cv1711.11586Adobe486 citesNov 30, 2017
- NovGraph Distillation for Action Detection with Privileged Modalitiesno summary yetcs-cv1712.00108Google Research3 citesNov 30, 2017
- OctLearning to Segment Human by Watching YouTubeno summary yetcs-cv1710.01457Snap24 citesOct 4, 2017
- OctGrader variability and the importance of reference standards for evaluating machine learning models for diabetic retinopathyno summary yetcs-cv1710.01711Google Research547 citesOct 4, 2017
- OctGeneralized Zero-Shot Learning for Action Recognition with Web-Scale Video Datano summary yetcs-cv1710.07455Tencent4 citesOct 20, 2017
- OctComplete 3D Scene Parsing from an RGBD Imageno summary yetcs-cv1710.09490Google Research1 citesOct 25, 2017
- OctDynamic Routing Between Capsulesno summary yetcs-cv1710.09829Google Research89 citesOct 26, 2017
- OctPoseTrack: A Benchmark for Human Pose Estimation and Trackingno summary yetcs-cv1710.10000Google Research32 citesOct 27, 2017
- OctDual Skipping Networksno summary yetcs-cv1710.10386Tencent0 citesOct 28, 2017
- SepHyperspectral Light Field Stereo Matchingno summary yetcs-cv1709.00835Microsoft Research0 citesSep 4, 2017
- SepTowards social pattern characterization in egocentric photo-streamsno summary yetcs-cv1709.01424Microsoft Research4 citesSep 5, 2017
- SepRobust Emotion Recognition from Low Quality and Low Bit Rate Video: A Deep Learning Approachno summary yetcs-cv1709.03126Snap6 citesSep 10, 2017
- SepEfficient Online Surface Correction for Real-time Large-Scale 3D Reconstructionno summary yetcs-cv1709.03763Google Research2 citesSep 12, 2017
- SepEmbedding Deep Networks into Visual Explanationsno summary yetcs-cv1709.05360Tencent23 citesSep 15, 2017
- SepNIMA: Neural Image Assessmentno summary yetcs-cv1709.05424Google Research933 citesSep 15, 2017
- SepJoint Detection and Recounting of Abnormal Events by Learning Deep Generic Knowledgeno summary yetcs-cv1709.09121Microsoft Research28 citesSep 26, 2017
- SepPhotorealistic Style Transfer with Screened Poisson Equationno summary yetcs-cv1709.09828Adobe9 citesSep 28, 2017
- AugLearning to Hallucinate Face Images via Component Generation and Enhancementno summary yetcs-cv1708.00223Tencent21 citesAug 1, 2017
- AugAssociative Domain Adaptationno summary yetcs-cv1708.00938Google Research39 citesAug 2, 2017
- AugLocalizing Moments in Video with Natural Languageno summary yetcs-cv1708.01641Adobe85 citesAug 4, 2017
- AugCut, Paste and Learn: Surprisingly Easy Synthesis for Instance Detectionno summary yetcs-cv1708.01642Google Research75 citesAug 4, 2017
- Aug3D-PRNN: Generating Shape Primitives with Recurrent Neural Networksno summary yetcs-cv1708.01648Adobe22 citesAug 4, 2017
- AugTraining Deep Networks to be Spatially Sensitiveno summary yetcs-cv1708.02212Adobe1 citesAug 7, 2017
- AugTips and Tricks for Visual Question Answering: Learnings from the 2017 Challengeno summary yetcs-cv1708.02711Microsoft Research53 citesAug 9, 2017
- AugExtreme clicking for efficient object annotationno summary yetcs-cv1708.02750Google Research20 citesAug 9, 2017
- AugAn evaluation of large-scale methods for image instance and class discoveryno summary yetcs-cv1708.02898Meta / FAIR1 citesAug 9, 2017
- AugSUBIC: A supervised, structured binary code for image searchno summary yetcs-cv1708.02932Amazon15 citesAug 9, 2017
- AugPersonalized Cinemagraphs using Semantic Understanding and Collaborative Learningno summary yetcs-cv1708.02970Microsoft Research3 citesAug 9, 2017
- AugDesnowNet: Context-Aware Deep Network for Snow Removalno summary yetcs-cv1708.04512Alibaba374 citesAug 15, 2017
- AugConvolutional Neural Networks for Non-iterative Reconstruction of Compressively Sensed Imagesno summary yetcs-cv1708.04669Apple10 citesAug 15, 2017
- AugGANs for Biological Image Synthesisno summary yetcs-cv1708.04692Amazon17 citesAug 15, 2017
- AugNavigator-free EPI Ghost Correction with Structured Low-Rank Matrix Models: New Theory and Methodsno summary yetcs-cv1708.05095IBM Research44 citesAug 16, 2017
- AugRevisiting knowledge transfer for training object class detectorsno summary yetcs-cv1708.06128Google Research7 citesAug 21, 2017
- AugCausally Regularized Learning with Agnostic Data Selection Biasno summary yetcs-cv1708.06656Tencent96 citesAug 22, 2017
- AugTags2Parts: Discovering Semantic Regions from Shape Tagsno summary yetcs-cv1708.06673Adobe2 citesAug 22, 2017
- AugGradient-based Camera Exposure Control for Outdoor Mobile Platformsno summary yetcs-cv1708.07338Adobe46 citesAug 24, 2017
- AugMulti-task Self-Supervised Visual Learningno summary yetcs-cv1708.07860DeepMind74 citesAug 25, 2017
- AugStylizing Face Images via Multiple Exemplarsno summary yetcs-cv1708.08288Tencent40 citesAug 28, 2017
- Aug3D Visual Perception for Self-Driving Cars using a Multi-Camera System: Calibration, Mapping, Localization, and Obstacle Detectionno summary yetcs-cv1708.09839Microsoft Research4 citesAug 31, 2017
- AugPredicting Cardiovascular Risk Factors from Retinal Fundus Photographs using Deep Learningno summary yetcs-cv1708.09843Google Research1,820 citesAug 31, 2017
- JulModulating early visual processing by languageno summary yetcs-cv1707.00683DeepMind298 citesJul 2, 2017
- JulEmbedding Visual Hierarchy with Deep Networks for Large-Scale Visual Recognitionno summary yetcs-cv1707.02406Amazon33 citesJul 8, 2017
- JulDiscriminative Optimization: Theory and Applications to Computer Vision Problemsno summary yetcs-cv1707.04318Meta / FAIR22 citesJul 13, 2017
- JulThe Reversible Residual Network: Backpropagation Without Storing Activationsno summary yetcs-cv1707.04585Google Research228 citesJul 14, 2017
- JulMoCoGAN: Decomposing Motion and Content for Video Generationno summary yetcs-cv1707.04993Snap104 citesJul 17, 2017
- JulShow and Recall: Learning What Makes Videos Memorableno summary yetcs-cv1707.05357Adobe14 citesJul 17, 2017
- JulHybrid PS-V Technique: A Novel Sensor Fusion Approach for Fast Mobile Eye-Tracking with Sensor-Shift Aware Correctionno summary yetcs-cv1707.05411Google Research0 citesJul 17, 2017
- JulPhotosensor Oculography: Survey and Parametric Analysis of Designs using Model-Based Simulationno summary yetcs-cv1707.05413Google Research1 citesJul 17, 2017
- JulDCTM: Discrete-Continuous Transformation Matching for Semantic Flowno summary yetcs-cv1707.05471Microsoft Research2 citesJul 18, 2017
- JulSkeleton-Based Human Action Recognition with Global Context-Aware Attention LSTM Networksno summary yetcs-cv1707.05740Alibaba523 citesJul 18, 2017
- JulThe Devil is in the Decoder: Classification, Regression and GANsno summary yetcs-cv1707.05847Google Research32 citesJul 18, 2017
- JulThe iNaturalist Species Classification and Detection Datasetno summary yetcs-cv1707.06642Google Research16 citesJul 20, 2017
- JulLearning Transferable Architectures for Scalable Image Recognitionno summary yetcs-cv1707.07012Google Research608 citesJul 21, 2017
- JulEyemotion: Classifying facial expressions in VR using eye-tracking camerasno summary yetcs-cv1707.07204Google Research18 citesJul 22, 2017
- JulDeep Co-Space: Sample Mining Across Feature Transformation for Semi-Supervised Learningno summary yetcs-cv1707.09119Tencent11 citesJul 28, 2017
- JunTextureGAN: Controlling Deep Image Synthesis with Texture Patchesno summary yetcs-cv1706.02823Adobe25 citesJun 9, 2017
- JunTeaching Compositionality to CNNsno summary yetcs-cv1706.04313Google Research8 citesJun 14, 2017
- MaySubmodular Trajectory Optimization for Aerial 3D Scanningno summary yetcs-cv1705.00703Adobe5 citesMay 1, 2017
- MayLesion detection and Grading of Diabetic Retinopathy via Two-stages Deep Convolutional Neural Networksno summary yetcs-cv1705.00771Baidu0 citesMay 2, 2017
- MayTemporal Segment Networks for Action Recognition in Videosno summary yetcs-cv1705.02953Amazon53 citesMay 8, 2017
- MayInferring and Executing Programs for Visual Reasoningno summary yetcs-cv1705.03633Meta / FAIR35 citesMay 10, 2017
- MayRe3 : Real-Time Recurrent Regression Networks for Visual Tracking of Generic Objectsno summary yetcs-cv1705.06368AllenAI9 citesMay 17, 2017
- MayExploring the structure of a real-time, arbitrary neural artistic stylization networkno summary yetcs-cv1705.06830Google Research38 citesMay 18, 2017
- MayPixColor: Pixel Recursive Colorizationno summary yetcs-cv1705.07208Google Research26 citesMay 19, 2017
- MayQuadruplet Network with One-Shot Learning for Fast Visual Object Trackingno summary yetcs-cv1705.07222Alibaba166 citesMay 19, 2017
- MayQuo Vadis, Action Recognition? A New Model and the Kinetics Datasetno summary yetcs-cv1705.07750DeepMind432 citesMay 22, 2017
- MayDepthCut: Improved Depth Edge Estimation Using Multiple Unreliable Channelsno summary yetcs-cv1705.07844Adobe0 citesMay 22, 2017
- MayUniversal Style Transfer via Feature Transformsno summary yetcs-cv1705.08086Adobe349 citesMay 23, 2017
- MayReinforced Temporal Attention and Split-Rate Transfer for Depth-Based Person Re-Identificationno summary yetcs-cv1705.09882Microsoft Research0 citesMay 28, 2017
- AprLearning to Predict Indoor Illumination from a Single Imageno summary yetcs-cv1704.00090Adobe13 citesApr 1, 2017
- AprThe Stixel world: A medium-level representation of traffic scenesno summary yetcs-cv1704.00280Microsoft Research43 citesApr 2, 2017
- AprActionVLAD: Learning spatio-temporal aggregation for action classificationno summary yetcs-cv1704.02895Adobe520 citesApr 10, 2017
- AprForecasting Human Dynamics from Static Imagesno summary yetcs-cv1704.03432Adobe14 citesApr 11, 2017
- AprAttention-based Extraction of Structured Information from Street View Imageryno summary yetcs-cv1704.03549Google Research31 citesApr 11, 2017
- AprDeep Reinforcement Learning-based Image Captioning with Embedding Rewardno summary yetcs-cv1704.03899Snap64 citesApr 12, 2017
- AprMulti-View Image Generation from a Single-Viewno summary yetcs-cv1704.04886Tencent41 citesApr 17, 2017
- AprDeep Self-Taught Learning for Weakly Supervised Object Localizationno summary yetcs-cv1704.05188Tencent28 citesApr 18, 2017
- AprIlluminant Spectra-based Source Separation Using Flash Photographyno summary yetcs-cv1704.05564Adobe1 citesApr 19, 2017
- AprLearning to Generate Long-term Future via Hierarchical Predictionno summary yetcs-cv1704.05831Adobe180 citesApr 19, 2017
- AprGenerative Face Completionno summary yetcs-cv1704.05838Adobe93 citesApr 19, 2017
- AprTraining object class detectors with click supervisionno summary yetcs-cv1704.06189Google Research15 citesApr 20, 2017
- AprTime-Contrastive Networks: Self-Supervised Learning from Videono summary yetcs-cv1704.06888Google Research16 citesApr 23, 2017
- AprSkeleton Key: Image Captioning by Skeleton-Attribute Decompositionno summary yetcs-cv1704.06972Adobe16 citesApr 23, 2017
- AprUnsupervised Learning of Depth and Ego-Motion from Videono summary yetcs-cv1704.07813Google Research222 citesApr 25, 2017
- AprICNet for Real-Time Semantic Segmentation on High-Resolution Imagesno summary yetcs-cv1704.08545Tencent130 citesApr 27, 2017
- AprBAM! The Behance Artistic Media Dataset for Recognition Beyond Photographyno summary yetcs-cv1704.08614Adobe20 citesApr 27, 2017
- AprA Unified Approach of Multi-scale Deep and Hand-crafted Features for Defocus Estimationno summary yetcs-cv1704.08992Tencent11 citesApr 28, 2017
- AprJoint Denoising / Compression of Image Contours via Shape Prior and Context Treeno summary yetcs-cv1705.00268Microsoft Research6 citesApr 30, 2017
- AprPredicting Foreground Object Ambiguity and Efficiently Crowdsourcing the Segmentation(s)no summary yetcs-cv1705.00366Adobe1 citesApr 30, 2017
- MarDiversified Texture Synthesis with Feed-forward Networksno summary yetcs-cv1703.01664Adobe26 citesMar 5, 2017
- MarTransformation-Grounded Image Generation Network for Novel 3D View Synthesisno summary yetcs-cv1703.02921Adobe37 citesMar 8, 2017
- MarInterpretable Structure-Evolving LSTMno summary yetcs-cv1703.03055Adobe14 citesMar 8, 2017
- MarDeep Image Mattingno summary yetcs-cv1703.03872Adobe33 citesMar 10, 2017
- MarProstate Cancer Diagnosis using Deep Learning with 3D Multiparametric MRIno summary yetcs-cv1703.04078LinkedIn88 citesMar 12, 2017
- MarLocal Patch Encoding-Based Method for Single Image Super-Resolutionno summary yetcs-cv1703.04088Snap0 citesMar 12, 2017
- MarConvolutional neural network architecture for geometric matchingno summary yetcs-cv1703.05593DeepMind513 citesMar 16, 2017
- MarRoomNet: End-to-End Room Layout Estimationno summary yetcs-cv1703.06241Google Research0 citesMar 18, 2017
- MarNo Fuss Distance Metric Learning using Proxiesno summary yetcs-cv1703.07464Google Research58 citesMar 21, 2017
- MarDeep Photo Style Transferno summary yetcs-cv1703.07511Adobe94 citesMar 22, 2017
- MarDiscriminative Transfer Learning for General Image Restorationno summary yetcs-cv1703.09245Meta / FAIR14 citesMar 27, 2017
- MarLucid Data Dreaming for Video Object Segmentationno summary yetcs-cv1703.09554Google Research24 citesMar 28, 2017
- MarImproved Lossy Image Compression with Priming and Spatially Adaptive Bit Rates for Recurrent Networksno summary yetcs-cv1703.10114Google Research27 citesMar 29, 2017
- FebPixel Recursive Super Resolutionno summary yetcs-cv1702.00783Google Research35 citesFeb 2, 2017
- FebYouTube-BoundingBoxes: A Large High-Precision Human-Annotated Data Set for Object Detection in Videono summary yetcs-cv1702.00824Google Research46 citesFeb 2, 2017
- FebDetailed Surface Geometry and Albedo Recovery from RGB-D Video Under Natural Illuminationno summary yetcs-cv1702.01486Baidu3 citesFeb 6, 2017
- FebCognitive Mapping and Planning for Visual Navigationno summary yetcs-cv1702.03920Google Research218 citesFeb 13, 2017
- FebEnd-to-End Interpretation of the French Street Name Signs Datasetno summary yetcs-cv1702.03970Google Research1 citesFeb 13, 2017
- FebEfficient Large-scale Approximate Nearest Neighbor Search on the GPUno summary yetcs-cv1702.05911Adobe58 citesFeb 20, 2017
- FebTransfer Learning for Domain Adaptation in MRI: Application in Brain Lesion Segmentationno summary yetcs-cv1702.07841IBM Research246 citesFeb 25, 2017
- JanLearning a Mixture of Deep Networks for Single Image Super-Resolutionno summary yetcs-cv1701.00823Adobe2 citesJan 3, 2017
- JanLearning From Noisy Large-Scale Datasets With Minimal Supervisionno summary yetcs-cv1701.01619Google Research54 citesJan 6, 2017
- JanTowards Accurate Multi-person Pose Estimation in the Wildno summary yetcs-cv1701.01779Google Research76 citesJan 6, 2017
- JanSee the Glass Half Full: Reasoning about Liquid Containers, their Volume and Contentno summary yetcs-cv1701.02718AllenAI7 citesJan 10, 2017
- JanA General and Adaptive Robust Loss Functionno summary yetcs-cv1701.03077Google Research10 citesJan 11, 2017
- JanComplex Event Recognition from Images with Few Training Examplesno summary yetcs-cv1701.04769Google Research4 citesJan 17, 2017
- JanSynthesizing Normalized Faces from Facial Identity Featuresno summary yetcs-cv1701.04851Google Research27 citesJan 17, 2017
- JanBringing Impressionism to Life with Neural Style Transfer in Come Swimno summary yetcs-cv1701.04928Adobe2 citesJan 18, 2017
- JanImage De-raining Using a Conditional Generative Adversarial Networkno summary yetcs-cv1701.05957Adobe249 citesJan 21, 2017
- JanA Survey of Structure from Motionno summary yetcs-cv1701.08493Meta / FAIR12 citesJan 30, 2017
2016
91- DecImproved Image Captioning via Policy Gradient optimization of SPIDErno summary yetcs-cv1612.00370Google Research454 citesDec 1, 2016
- DecPerspective Transformer Nets: Learning Single-View 3D Object Reconstruction without 3D Supervisionno summary yetcs-cv1612.00814Adobe316 citesDec 1, 2016
- DecScribbler: Controlling Deep Image Synthesis with Sketch and Colorno summary yetcs-cv1612.00835Adobe26 citesDec 2, 2016
- DecMaking the V in VQA Matter: Elevating the Role of Image Understanding in Visual Question Answeringno summary yetcs-cv1612.00837Meta / FAIR40 citesDec 2, 2016
- DecDeep Metric Learning via Facility Locationno summary yetcs-cv1612.01213Google Research6 citesDec 5, 2016
- DecLearning to Detect Multiple Photographic Defectsno summary yetcs-cv1612.01635Adobe3 citesDec 6, 2016
- DecKnowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioningno summary yetcs-cv1612.01887Salesforce54 citesDec 6, 2016
- DecTag Prediction at Flickr: a View from the Darkroomno summary yetcs-cv1612.01922Meta / FAIR6 citesDec 6, 2016
- DecSaliency Driven Image Manipulationno summary yetcs-cv1612.02184Adobe4 citesDec 7, 2016
- DecSpatially Adaptive Computation Time for Residual Networksno summary yetcs-cv1612.02297Google Research25 citesDec 7, 2016
- DecStackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networksno summary yetcs-cv1612.03242Baidu706 citesDec 10, 2016
- DecUnsupervised Pixel-Level Domain Adaptation with Generative Adversarial Networksno summary yetcs-cv1612.05424Google Research72 citesDec 16, 2016
- DecLearning Features by Watching Objects Moveno summary yetcs-cv1612.06370Meta / FAIR18 citesDec 19, 2016
- DecAsynchronous Temporal Fields for Action Recognitionno summary yetcs-cv1612.06371AllenAI9 citesDec 19, 2016
- DecUnsupervised Perceptual Rewards for Imitation Learningno summary yetcs-cv1612.06699Google Research5 citesDec 20, 2016
- DecA Statistical Approach to Continuous Self-Calibrating Eye Gaze Tracking for Head-Mounted Virtual Reality Systemsno summary yetcs-cv1612.06919Microsoft Research0 citesDec 20, 2016
- DecTemporal Tessellation: A Unified Approach for Video Analysisno summary yetcs-cv1612.06950Amazon4 citesDec 21, 2016
- DecTop-down Visual Saliency Guided by Captionsno summary yetcs-cv1612.07360Adobe6 citesDec 21, 2016
- DecLearning Visual N-Grams from Web Datano summary yetcs-cv1612.09161Meta / FAIR0 citesDec 29, 2016
- NovAdversarial Machine Learning at Scaleno summary yetcs-cv1611.01236Google Research377 citesNov 4, 2016
- NovVariational Deep Embedding: An Unsupervised and Generative Approach to Clusteringno summary yetcs-cv1611.05148Tencent89 citesNov 16, 2016
- NovSCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioningno summary yetcs-cv1611.05594Tencent39 citesNov 17, 2016
- NovDeep Outdoor Illumination Estimationno summary yetcs-cv1611.06403Adobe13 citesNov 19, 2016
- NovRecurrent Memory Addressing for describing videosno summary yetcs-cv1611.06492Google Research2 citesNov 20, 2016
- NovExploiting Web Images for Dataset Construction: A Domain Robust Approachno summary yetcs-cv1611.07156Alibaba76 citesNov 22, 2016
- NovFast Fourier Color Constancyno summary yetcs-cv1611.07596Google Research3 citesNov 23, 2016
- NovMulti-View 3D Object Detection Network for Autonomous Drivingno summary yetcs-cv1611.07759Baidu66 citesNov 23, 2016
- NovControlling Perceptual Factors in Neural Style Transferno summary yetcs-cv1611.07865Adobe11 citesNov 23, 2016
- NovGeometric deep learning: going beyond Euclidean datano summary yetcs-cv1611.08097Meta / FAIR3,594 citesNov 24, 2016
- NovLearning a Discriminative Filter Bank within a CNN for Fine-grained Recognitionno summary yetcs-cv1611.09932Adobe9 citesNov 29, 2016
- NovSpeed/accuracy trade-offs for modern convolutional object detectorsno summary yetcs-cv1611.10012Google Research157 citesNov 30, 2016
- NovMultimodal Transfer: A Hierarchical Deep Convolutional Neural Network for Fast Artistic Style Transferno summary yetcs-cv1612.01895Adobe10 citesNov 17, 2016
- OctVideo Pixel Networksno summary yetcs-cv1610.00527Google Research22 citesOct 3, 2016
- OctXception: Deep Learning with Depthwise Separable Convolutionsno summary yetcs-cv1610.02357Google Research358 citesOct 7, 2016
- OctGrad-CAM: Visual Explanations from Deep Networks via Gradient-based Localizationno summary yetcs-cv1610.02391Meta / FAIR5,501 citesOct 7, 2016
- OctA Learned Representation For Artistic Styleno summary yetcs-cv1610.07629Google Research171 citesOct 24, 2016
- SepBest-Buddies Similarity - Robust Template Matching using Mutual Nearest Neighborsno summary yetcs-cv1609.01571Google Research5 citesSep 6, 2016
- SepStyle-Transfer via Texture-Synthesisno summary yetcs-cv1609.03057Google Research159 citesSep 10, 2016
- SepA Tube-and-Droplet-based Approach for Representing and Analyzing Motion Trajectoriesno summary yetcs-cv1609.03058Microsoft Research58 citesSep 10, 2016
- SepSemi-Supervised Sparse Representation Based Classification for Face Recognition with Insufficient Labeled Samplesno summary yetcs-cv1609.03279Tencent255 citesSep 12, 2016
- SepActive Canny: Edge Detection and Recovery with Open Active Contour Modelsno summary yetcs-cv1609.03415NVIDIA1 citesSep 12, 2016
- SepGenerative Visual Manipulation on the Natural Image Manifoldno summary yetcs-cv1609.03552Adobe122 citesSep 12, 2016
- SepSingle-image RGB Photometric Stereo With Spatially-varying Albedono summary yetcs-cv1609.04079Adobe3 citesSep 14, 2016
- SepDeep CTR Prediction in Display Advertisingno summary yetcs-cv1609.06018Alibaba0 citesSep 20, 2016
- SepShow and Tell: Lessons learned from the 2015 MSCOCO Image Captioning Challengeno summary yetcs-cv1609.06647Google Research912 citesSep 21, 2016
- SepLinear Support Tensor Machine: Pedestrian Detection in Thermal Infrared Imagesno summary yetcs-cv1609.07878Google Research72 citesSep 26, 2016
- SepMulti-view Self-supervised Deep Learning for 6D Pose Estimation in the Amazon Picking Challengeno summary yetcs-cv1609.09475Google Research36 citesSep 29, 2016
- AugTraining Deep Networks for Facial Expression Recognition with Crowd-Sourced Label Distributionno summary yetcs-cv1608.01041Microsoft Research11 citesAug 3, 2016
- AugLearning Common and Specific Features for RGB-D Semantic Segmentation with Deconvolutional Networksno summary yetcs-cv1608.01082NVIDIA12 citesAug 3, 2016
- AugDesign of Efficient Convolutional Layers using Single Intra-channel Convolution, Topological Subdivisioning and Spatial "Bottleneck" Structureno summary yetcs-cv1608.04337Amazon27 citesAug 15, 2016
- AugIntrinsic Light Field Imagesno summary yetcs-cv1608.04342Adobe13 citesAug 15, 2016
- AugGlobally Variance-Constrained Sparse Representation and Its Application in Image Set Codingno summary yetcs-cv1608.04902Tencent11 citesAug 17, 2016
- AugFull Resolution Image Compression with Recurrent Neural Networksno summary yetcs-cv1608.05148Google Research71 citesAug 18, 2016
- AugA Recurrent Encoder-Decoder Network for Sequential Face Alignmentno summary yetcs-cv1608.05477Snap33 citesAug 19, 2016
- AugDomain Separation Networksno summary yetcs-cv1608.06019Google Research767 citesAug 22, 2016
- AugAmbient Sound Provides Supervision for Visual Learningno summary yetcs-cv1608.07017Google Research0 citesAug 25, 2016
- AugVehicle Detection from 3D Lidar Using Fully Convolutional Networkno summary yetcs-cv1608.07916Baidu275 citesAug 29, 2016
- AugMulti-Class Multi-Object Tracking using Changing Point Detectionno summary yetcs-cv1608.08434Meta / FAIR77 citesAug 30, 2016
- JulUnsupervised Learning of 3D Structure from Imagesno summary yetcs-cv1607.00662Google Research98 citesJul 3, 2016
- JulVisual Dynamics: Probabilistic Future Frame Synthesis via Cross Convolutional Networksno summary yetcs-cv1607.02586Google Research145 citesJul 9, 2016
- JulAccelerating Eulerian Fluid Simulation With Convolutional Networksno summary yetcs-cv1607.03597Google Research296 citesJul 13, 2016
- JulDSD: Dense-Sparse-Dense Training for Deep Neural Networksno summary yetcs-cv1607.04381Baidu29 citesJul 15, 2016
- JulExploiting Symmetry and/or Manhattan Properties for 3D Object Structure Estimation from Single and Multiple Imagesno summary yetcs-cv1607.07129Tencent6 citesJul 25, 2016
- JunMultiview Rectification of Folded Documentsno summary yetcs-cv1606.00166Microsoft Research1 citesJun 1, 2016
- JunRAISR: Rapid and Accurate Image Super Resolutionno summary yetcs-cv1606.01299Google Research6 citesJun 3, 2016
- JunPhoto Aesthetics Ranking Network with Attributes and Content Adaptationno summary yetcs-cv1606.01621Adobe40 citesJun 6, 2016
- JunHuman Attention in Visual Question Answering: Do Humans and Deep Networks Look at the Same Regions?no summary yetcs-cv1606.03556Meta / FAIR51 citesJun 11, 2016
- JunConditional Image Generation with PixelCNN Decodersno summary yetcs-cv1606.05328DeepMind1,589 citesJun 16, 2016
- Jun3D U-Net: Learning Dense Volumetric Segmentation from Sparse Annotationno summary yetcs-cv1606.06650DeepMind494 citesJun 21, 2016
- JunFast Multi-Layer Laplacian Enhancementno summary yetcs-cv1606.07396Google Research0 citesJun 23, 2016
- MayR-FCN: Object Detection via Region-based Fully Convolutional Networksno summary yetcs-cv1605.06409Microsoft Research3,438 citesMay 20, 2016
- AprMarr Revisited: 2D-3D Alignment via Surface Normal Predictionno summary yetcs-cv1604.01347Adobe34 citesApr 5, 2016
- AprSubcategory-aware Convolutional Neural Networks for Object Proposals and Detectionno summary yetcs-cv1604.04693Baidu15 citesApr 16, 2016
- AprThe THUMOS Challenge on Action Recognition for Videos "in the Wild"no summary yetcs-cv1604.06182Google Research588 citesApr 21, 2016
- AprVisual Congruent Ads for Image Searchno summary yetcs-cv1604.06481Amazon0 citesApr 21, 2016
- AprText Flow: A Unified Text Detection System in Natural Scene Imagesno summary yetcs-cv1604.06877Baidu1 citesApr 23, 2016
- AprDeep Edge Guided Recurrent Residual Learning for Image Super-Resolutionno summary yetcs-cv1604.08671Snap206 citesApr 29, 2016
- AprSingle Image 3D Interpreter Networkno summary yetcs-cv1604.08685Meta / FAIR95 citesApr 29, 2016
- MarConvolutional Patch Representations for Image Retrieval: an Unsupervised Approachno summary yetcs-cv1603.00438Meta / FAIR5 citesMar 1, 2016
- MarDrift Robust Non-rigid Optical Flow Enhancement for Long Sequencesno summary yetcs-cv1603.02252Google Research0 citesMar 7, 2016
- MarDeep Interactive Object Selectionno summary yetcs-cv1603.04042Adobe42 citesMar 13, 2016
- MarA Diagram Is Worth A Dozen Imagesno summary yetcs-cv1603.07396AllenAI13 citesMar 24, 2016
- MarFine-scale Surface Normal Estimation using a Single NIR Imageno summary yetcs-cv1603.07475Adobe0 citesMar 24, 2016
- MarAttend, Infer, Repeat: Fast Scene Understanding with Generative Modelsno summary yetcs-cv1603.08575DeepMind250 citesMar 28, 2016
- MarPartial Face Detection for Continuous Authenticationno summary yetcs-cv1603.09364Google Research48 citesMar 30, 2016
- FebPlaNet - Photo Geolocation with Convolutional Neural Networksno summary yetcs-cv1602.05314Google Research392 citesFeb 17, 2016
- FebInception-v4, Inception-ResNet and the Impact of Residual Connections on Learningno summary yetcs-cv1602.07261Google Research4,493 citesFeb 23, 2016
- JanJoint Object-Material Category Segmentation from Audio-Visual Cuesno summary yetcs-cv1601.02220Microsoft Research1 citesJan 10, 2016
- JanPixel Recurrent Neural Networksno summary yetcs-cv1601.06759DeepMind1,312 citesJan 25, 2016
- JanDeep Learning Driven Visual Path Prediction from a Single Imageno summary yetcs-cv1601.07265Tencent65 citesJan 27, 2016
- JanA Grassmannian Graph Approach to Affine Invariant Feature Matchingno summary yetcs-cv1601.07648Microsoft Research1 citesJan 28, 2016
2015
64- DecLabeling the Features Not the Samples: Efficient Video Classification with Minimal Supervisionno summary yetcs-cv1512.00517Google Research6 citesDec 1, 2015
- DecASIST: Automatic Semantically Invariant Scene Transformationno summary yetcs-cv1512.01515Microsoft Research0 citesDec 4, 2015
- DecSSD: Single Shot MultiBox Detectorno summary yetcs-cv1512.02325Google Research20,854 citesDec 8, 2015
- DecDeep Exemplar 2D-3D Detection by Adapting from Real to Rendered Viewsno summary yetcs-cv1512.02497Adobe17 citesDec 8, 2015
- DecWe Are Humor Beings: Understanding and Predicting Visual Humorno summary yetcs-cv1512.04407Meta / FAIR5 citesDec 14, 2015
- DecBlockout: Dynamic Model Selection for Hierarchical Deep Networksno summary yetcs-cv1512.05246Google Research22 citesDec 16, 2015
- DecWrite a Classifier: Predicting Visual Classifiers from Unstructured Textno summary yetcs-cv1601.00025Meta / FAIR3 citesDec 31, 2015
- Nov3D Time-lapse Reconstruction from Internet Photosno summary yetcs-cv1511.03019Google Research1 citesNov 10, 2015
- NovThe Fast Bilateral Solverno summary yetcs-cv1511.03296Google Research9 citesNov 10, 2015
- NovSemantic Image Segmentation with Task-Specific Edge Detection Using CNNs and a Discriminatively Trained Domain Transformno summary yetcs-cv1511.03328Google Research35 citesNov 10, 2015
- NovAutomatic Content-Aware Color and Tone Stylizationno summary yetcs-cv1511.03748Adobe3 citesNov 12, 2015
- NovNewtonian Image Understanding: Unfolding the Dynamics of Objects in Static Imagesno summary yetcs-cv1511.04048AllenAI21 citesNov 12, 2015
- NovUnsupervised Learning of Edgesno summary yetcs-cv1511.04166Meta / FAIR7 citesNov 13, 2015
- NovSemantic Object Parsing with Local-Global Long Short-Term Memoryno summary yetcs-cv1511.04510Adobe30 citesNov 14, 2015
- NovReversible Recursive Instance-level Object Segmentationno summary yetcs-cv1511.04517Adobe11 citesNov 14, 2015
- NovLearning Fine-grained Features via a CNN Tree for Large-scale Classificationno summary yetcs-cv1511.04534Alibaba7 citesNov 14, 2015
- NovSherlock: Scalable Fact Learning in Imagesno summary yetcs-cv1511.04891Adobe6 citesNov 16, 2015
- NovDense Human Body Correspondences Using Convolutional Networksno summary yetcs-cv1511.05904Adobe13 citesNov 18, 2015
- NovCompact Bilinear Poolingno summary yetcs-cv1511.06062Snap46 citesNov 19, 2015
- NovVariable Rate Image Compression with Recurrent Neural Networksno summary yetcs-cv1511.06085Google Research118 citesNov 19, 2015
- NovThe Unreasonable Effectiveness of Noisy Data for Fine-Grained Recognitionno summary yetcs-cv1511.06789Google Research20 citesNov 20, 2015
- NovBehavior Discovery and Alignment of Articulated Object Classes from Unstructured Videono summary yetcs-cv1511.09319Google Research18 citesNov 30, 2015
- OctOn the Existence of Epipolar Matricesno summary yetcs-cv1510.01401Google Research3 citesOct 6, 2015
- OctEgocentric Field-of-View Localization Using First-Person Point-of-View Devicesno summary yetcs-cv1510.02073Google Research28 citesOct 7, 2015
- OctSemanticPaint: A Framework for the Interactive Segmentation of 3D Scenesno summary yetcs-cv1510.03727Microsoft Research77 citesOct 13, 2015
- OctContent adaptive screen image scalingno summary yetcs-cv1510.06093Microsoft Research0 citesOct 21, 2015
- OctFinding Temporally Consistent Occlusion Boundaries in Videos using Geometric Contextno summary yetcs-cv1510.07323NVIDIA5 citesOct 25, 2015
- OctVideo Paragraph Captioning Using Hierarchical Recurrent Neural Networksno summary yetcs-cv1510.07712Baidu86 citesOct 26, 2015
- OctVISALOGY: Answering Visual Analogy Questionsno summary yetcs-cv1510.08973Microsoft Research21 citesOct 30, 2015
- SepFast Randomized Singular Value Thresholding for Low-rank Optimizationno summary yetcs-cv1509.00296Tencent1 citesSep 1, 2015
- SepSTC: A Simple to Complex Framework for Weakly-supervised Semantic Segmentationno summary yetcs-cv1509.03150Adobe615 citesSep 10, 2015
- SepA deep matrix factorization method for learning attribute representationsno summary yetcs-cv1509.03248Google Research4 citesSep 10, 2015
- SepRobust Image Sentiment Analysis Using Progressively Trained and Domain Transferred Deep Networksno summary yetcs-cv1509.06041Adobe213 citesSep 20, 2015
- SepVision System and Depth Processing for DRC-HUBO+no summary yetcs-cv1509.06114Adobe0 citesSep 21, 2015
- JulConvolutional Color Constancyno summary yetcs-cv1507.00410Google Research21 citesJul 2, 2015
- JulFace Alignment Assisted by Head Pose Estimationno summary yetcs-cv1507.03148Alibaba18 citesJul 11, 2015
- JulDeepFont: Identify Your Font from An Imageno summary yetcs-cv1507.03196Snap9 citesJul 12, 2015
- JulHuman Pose Estimation with Iterative Error Feedbackno summary yetcs-cv1507.06550DeepMind119 citesJul 23, 2015
- JunUnderstanding deep features with computer-generated imageryno summary yetcs-cv1506.01151Adobe34 citesJun 3, 2015
- JunFlowing ConvNets for Human Pose Estimation in Videosno summary yetcs-cv1506.02897Google Research90 citesJun 9, 2015
- JunLearning to Linearize Under Uncertaintyno summary yetcs-cv1506.03011Meta / FAIR34 citesJun 9, 2015
- JunCombinatorial Energy Learning for Image Segmentationno summary yetcs-cv1506.04304Google Research10 citesJun 13, 2015
- JunDeep Generative Image Models using a Laplacian Pyramid of Adversarial Networksno summary yetcs-cv1506.05751Meta / FAIR1,660 citesJun 18, 2015
- JunGraph-based compression of dynamic 3D point cloud sequencesno summary yetcs-cv1506.06096Microsoft Research223 citesJun 19, 2015
- JunGeneralized Majorization-Minimizationno summary yetcs-cv1506.07613Google Research9 citesJun 25, 2015
- MayJoint Object and Part Segmentation using Deep Learned Potentialsno summary yetcs-cv1505.00276Adobe27 citesMay 1, 2015
- MayContextual Action Recognition with R*CNNno summary yetcs-cv1505.01197Microsoft Research72 citesMay 5, 2015
- MayAutomatic Script Identification in the Wildno summary yetcs-cv1505.02982Tencent4 citesMay 12, 2015
- MayAre You Talking to a Machine? Dataset and Methods for Multilingual Image Question Answeringno summary yetcs-cv1505.05612Baidu68 citesMay 21, 2015
- MayInner and Inter Label Propagation: Salient Object Detection in the Wildno summary yetcs-cv1505.07192Adobe205 citesMay 27, 2015
- AprTemporal Localization of Fine-Grained Actions in Videos by Domain Transfer from Web Imagesno summary yetcs-cv1504.00983Google Research127 citesApr 4, 2015
- AprLocally Non-rigid Registration for Mobile HDR Photographyno summary yetcs-cv1504.01441NVIDIA2 citesApr 7, 2015
- AprA robust and efficient video representation for action recognitionno summary yetcs-cv1504.05524Amazon17 citesApr 21, 2015
- AprPerforatedCNNs: Acceleration through Elimination of Redundant Convolutionsno summary yetcs-cv1504.08362Microsoft Research119 citesApr 30, 2015
- MarLearning Super-Resolution Jointly from External and Internal Examplesno summary yetcs-cv1503.01138Snap77 citesMar 3, 2015
- MarDeep Human Parsing with Active Template Regressionno summary yetcs-cv1503.02391Adobe327 citesMar 9, 2015
- MarDeep Convolutional Inverse Graphics Networkno summary yetcs-cv1503.03167Microsoft Research752 citesMar 11, 2015
- MarDesigning A Composite Dictionary Adaptively From Joint Examplesno summary yetcs-cv1503.03621Snap2 citesMar 12, 2015
- MarBeyond Short Snippets: Deep Networks for Video Classificationno summary yetcs-cv1503.08909Google Research256 citesMar 31, 2015
- MarReal-World Font Recognition Using Deep Network and Domain Adaptationno summary yetcs-cv1504.00028Adobe5 citesMar 31, 2015
- FebConditional Random Fields as Recurrent Neural Networksno summary yetcs-cv1502.03240Baidu2,417 citesFeb 11, 2015
- FebWhat makes for effective detection proposals?no summary yetcs-cv1502.05082Meta / FAIR757 citesFeb 17, 2015
- FebVIP: Finding Important People in Imagesno summary yetcs-cv1502.05678Google Research0 citesFeb 19, 2015
- FebCoercive Region-level Registration for Multi-modal Imagesno summary yetcs-cv1502.07432Google Research3 citesFeb 26, 2015
2014
25- DecConvolutional Feature Masking for Joint Object and Stuff Segmentationno summary yetcs-cv1412.1283Microsoft Research470 citesDec 3, 2014
- DecConvolutional Neural Networks at Constrained Time Costno summary yetcs-cv1412.1710Microsoft Research27 citesDec 4, 2014
- DecActions and Attributes from Wholes and Partsno summary yetcs-cv1412.2604Microsoft Research11 citesDec 8, 2014
- DecDeep Structured Output Learning for Unconstrained Text Recognitionno summary yetcs-cv1412.5903DeepMind164 citesDec 18, 2014
- DecTraining Deep Neural Networks on Noisy Labels with Bootstrappingno summary yetcs-cv1412.6596Microsoft Research330 citesDec 20, 2014
- DecDeep Captioning with Multimodal Recurrent Neural Networks (m-RNN)no summary yetcs-cv1412.6632Baidu653 citesDec 20, 2014
- DecAttention for Fine-Grained Categorizationno summary yetcs-cv1412.7054Google Research119 citesDec 22, 2014
- DecSemantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFsno summary yetcs-cv1412.7062Google Research3,640 citesDec 22, 2014
- DecAutomatic Photo Adjustment Using Deep Neural Networksno summary yetcs-cv1412.7725Adobe4 citesDec 24, 2014
- NovAnisotropic Agglomerative Adaptive Mean-Shiftno summary yetcs-cv1411.4102Google Research0 citesNov 15, 2014
- NovEfficient and Accurate Approximations of Nonlinear Convolutional Networksno summary yetcs-cv1411.4229Microsoft Research3 citesNov 16, 2014
- NovCIDEr: Consensus-based Image Description Evaluationno summary yetcs-cv1411.5726Microsoft Research61 citesNov 20, 2014
- NovHypercolumns for Object Segmentation and Fine-grained Localizationno summary yetcs-cv1411.5752Microsoft Research21 citesNov 21, 2014
- NovAn Egocentric Look at Video Photographer Identityno summary yetcs-cv1411.7591Google Research4 citesNov 27, 2014
- NovArticulated motion discovery using pairs of trajectoriesno summary yetcs-cv1411.7883Google Research1 citesNov 28, 2014
- OctCapturing spatial interdependence in image features: the counting grid, an epitomic representation for bags of featuresno summary yetcs-cv1410.6264Microsoft Research0 citesOct 23, 2014
- OctConsensus Message Passing for Layered Graphical Modelsno summary yetcs-cv1410.7452Microsoft Research3 citesOct 27, 2014
- SepDeformable Part Models are Convolutional Neural Networksno summary yetcs-cv1409.5403Microsoft Research58 citesSep 18, 2014
- JunWeb-Scale Training for Face Identificationno summary yetcs-cv1406.5266Meta / FAIR13 citesJun 20, 2014
- JunFast Edge Detection Using Structured Forestsno summary yetcs-cv1406.5549Microsoft Research18 citesJun 20, 2014
- AprCascades of Regression Tree Fields for Image Restorationno summary yetcs-cv1404.2086Microsoft Research89 citesApr 8, 2014
- AprScalable Similarity Learning using Large Margin Neighborhood Embeddingno summary yetcs-cv1404.6272Adobe0 citesApr 24, 2014
- MarA Novel Method to Extract Rocks from Mars Imagesno summary yetcs-cv1403.3083Tencent3 citesMar 13, 2014
- FebFine-Grained Visual Categorization via Multi-stage Metric Learningno summary yetcs-cv1402.0453Alibaba16 citesFeb 3, 2014
- JanGeneralized Bhattacharyya and Chernoff upper bounds on Bayes error using quasi-arithmetic meansno summary yetcs-cv1401.4788Sony49 citesJan 20, 2014
2013
10- DecScalable Object Detection using Deep Neural Networksno summary yetcs-cv1312.2249Google Research36 citesDec 8, 2013
- DecDeepPose: Human Pose Estimation via Deep Neural Networksno summary yetcs-cv1312.4659Google Research3,244 citesDec 17, 2013
- DecLearning High-level Image Representation for Image Retrieval via Multi-Task DNN using Clickthrough Datano summary yetcs-cv1312.4740Microsoft Research1 citesDec 17, 2013
- DecDeep Convolutional Ranking for Multilabel Image Annotationno summary yetcs-cv1312.4894Google Research258 citesDec 17, 2013
- DecMulti-digit Number Recognition from Street View Imagery using Deep Convolutional Neural Networksno summary yetcs-cv1312.6082Google Research437 citesDec 20, 2013
- DecGPU Asynchronous Stochastic Gradient Descent to Speed Up Neural Network Trainingno summary yetcs-cv1312.6186Adobe67 citesDec 21, 2013
- DecOne-Shot Adaptation of Supervised Deep Convolutional Modelsno summary yetcs-cv1312.6204Google Research15 citesDec 21, 2013
- MarGBM Volumetry using the 3D Slicer Medical Image Computing Platformno summary yetcs-cv1303.0964IBM Research256 citesMar 5, 2013
- JanComplexity of Representation and Inference in Compositional Models with Part Sharingno summary yetcs-cv1301.3560AllenAI1 citesJan 16, 2013
- JanRegularized Discriminant Embedding for Visual Descriptor Learningno summary yetcs-cv1301.3644Microsoft Research0 citesJan 16, 2013
2012
5- DecPituitary Adenoma Volumetry with 3D Slicerno summary yetcs-cv1212.2860IBM Research92 citesDec 12, 2012
- JulProbabilistic index maps for modeling natural signalsno summary yetcs-cv1207.4179Microsoft Research0 citesJul 12, 2012
- MarSquare-Cut: A Segmentation Algorithm on the Basis of a Rectangle Shapeno summary yetcs-cv1203.2839IBM Research60 citesMar 13, 2012
- MarReal-time Image-based 6-DOF Localization in Large-Scale Environmentsno summary yetcs-cv1203.4355Microsoft Research18 citesMar 20, 2012
- FebGeneralized Boundaries from Multiple Image Interpretationsno summary yetcs-cv1202.3684Google Research3 citesFeb 16, 2012