| Locality Preserving Matching |
81 |
| Deep Expectation of Real and Apparent Age from a Single Image Without Facial Landmarks |
57 |
| Semantic Understanding of Scenes Through the ADE20K Dataset |
46 |
| From BoW to CNN: Two Decades of Texture Representation for Texture Classification |
45 |
| Facial Landmark Detection: A Literature Survey |
29 |
| Discriminative Correlation Filter Tracker with Channel and Spatial Reliability |
27 |
| Video Enhancement with Task-Oriented Flow |
27 |
| Top-Down Neural Attention by Excitation Backprop |
27 |
| On Unifying Multi-view Self-Representations for Clustering by Tensor Multi-rank Minimization |
26 |
| Augmented Reality Meets Computer Vision: Efficient Data Generation for Urban Driving Scenes |
26 |
| Beyond Temporal Pooling: Recurrence and Temporal Convolutions for Gesture Recognition in Video |
26 |
| Semantic Foggy Scene Understanding with Synthetic Data |
25 |
| Every Moment Counts: Dense Detailed Labeling of Actions in Complex Videos |
24 |
| Large Scale 3D Morphable Models |
23 |
| Adaptive Correlation Filters with Long-Term and Short-Term Memory for Object Tracking |
21 |
| From Facial Expression Recognition to Interpersonal Relation Prediction |
18 |
| Leveraging Prior-Knowledge for Weakly Supervised Object Detection Under a Collaborative Self-Paced Curriculum Learning Framework |
18 |
| Cluster Sparsity Field: An Internal Hyperspectral Imagery Prior for Reconstruction |
16 |
| Deep Sign: Enabling Robust Statistical Continuous Sign Language Recognition via Hybrid CNN-HMMs |
15 |
| Blind Image Deblurring via Deep Discriminative Priors |
14 |
| Hierarchical Cellular Automata for Visual Saliency |
14 |
| A Comprehensive Performance Evaluation of Deformable Face Tracking In-the-Wild |
14 |
| What Makes Good Synthetic Training Data for Learning Disparity and Optical Flow Estimation? |
12 |
| Context-Based Path Prediction for Targets with Switching Dynamics |
11 |
| Real-Time Intensity-Image Reconstruction for Event Cameras Using Manifold Regularisation |
11 |
| Which and How Many Regions to Gaze: Focus Discriminative Regions for Fine-Grained Visual Categorization |
10 |
| Large-Scale Bisample Learning on ID Versus Spot Face Recognition |
10 |
| The Menpo Benchmark for Multi-pose 2D and 3D Facial Landmark Localisation and Tracking |
10 |
| Multi-label Learning with Missing Labels Using Mixed Dependency Graphs |
10 |
| Zoom Out-and-In Network with Map Attention Decision for Region Proposal and Object Detection |
10 |
| Sim4CV: A Photo-Realistic Simulator for Computer Vision Applications |
9 |
| Learning Latent Representations of 3D Human Pose with Deep Neural Networks |
9 |
| Fusing Visual and Inertial Sensors with Semantics for 3D Human Pose Estimation |
9 |
| Artistic Style Transfer for Videos and Spherical Images |
9 |
| EMVS: Event-Based Multi-View Stereo3D Reconstruction with an Event Camera in Real-Time |
8 |
| Depth-Based Hand Pose Estimation: Methods, Data, and Challenges |
8 |
| Attentive Systems: A Survey |
8 |
| Deep Affect Prediction in-the-Wild: Aff-Wild Database and Challenge, Deep Architectures, and Beyond |
8 |
| Learning to Segment Moving Objects |
7 |
| Transferring Deep Object and Scene Representations for Event Recognition in Still Images |
7 |
| A Comprehensive Study on Center Loss for Deep Face Recognition |
7 |
| Learning Discriminative Aggregation Network for Video-Based Face Recognition and Person Re-identification |
7 |
| Do Semantic Parts Emerge in Convolutional Neural Networks? |
7 |
| Visual Tracking via Subspace Learning: A Discriminative Approach |
7 |
| Person Re-identification in Identity Regression Space |
6 |
| Hierarchical Attention for Part-Aware Face Detection |
6 |
| Baseline and Triangulation Geometry in a Standard Plenoptic Camera |
6 |
| Joint Face Hallucination and Deblurring via Structure Generation and Detail Enhancement |
6 |
| Unsupervised Binary Representation Learning with Deep Variational Networks |
6 |
| End-to-End Learning of Latent Deformable Part-Based Representations for Object Detection |
5 |