Scene Understanding

Scene understanding involves interpreting the visual information of a scene, including objects, their spatial relationships, and the overall layout. It goes beyond simple object recognition by considering the context and how objects relate to each other and the environment.

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 451–475 of 1723 papers

Title	Date	Tasks	Status	Hype
Image Segmentation Using Deep Learning: A Survey	Jan 15, 2020	DecoderDeep Learning	CodeCode Available	1
NODIS: Neural Ordinary Differential Scene Understanding	Jan 14, 2020	AllGraph Generation	CodeCode Available	1
Visual-Semantic Graph Attention Networks for Human-Object Interaction Detection	Jan 7, 2020	Graph AttentionHuman-Object Interaction Detection	CodeCode Available	1
IRS: A Large Naturalistic Indoor Robotics Stereo Dataset to Train Deep Models for Disparity and Surface Normal Estimation	Dec 20, 2019	Disparity EstimationScene Understanding	CodeCode Available	1
AeroRIT: A New Scene for Hyperspectral Image Analysis	Dec 17, 2019	Hyperspectral image analysisImage Super-Resolution	CodeCode Available	1
TextSLAM: Visual SLAM with Planar Text Features	Nov 26, 2019	Object SLAMScene Understanding	CodeCode Available	1
Towards Ghost-free Shadow Removal via Dual Hierarchical Aggregation Network and Shadow Matting GAN	Nov 20, 2019	2kGenerative Adversarial Network	CodeCode Available	1
DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames	Nov 1, 2019	Autonomous NavigationGPU	CodeCode Available	1
Underwater Image Super-Resolution using Deep Residual Multipliers	Sep 20, 2019	Image Super-ResolutionScene Understanding	CodeCode Available	1
Global Aggregation then Local Distribution in Fully Convolutional Networks	Sep 16, 2019	Instance Segmentationobject-detection	CodeCode Available	1
Dynamic Graph Message Passing Networks	Aug 19, 2019	Image Classificationobject-detection	CodeCode Available	1
VideoNavQA: Bridging the Gap between Visual and Embodied Question Answering	Aug 14, 2019	Embodied Question AnsweringQuestion Answering	CodeCode Available	1
M3D-RPN: Monocular 3D Region Proposal Network for Object Detection	Jul 13, 2019	3D Object Detection3D Object Detection From Monocular Images	CodeCode Available	1
From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network	Jul 8, 2019	3D Object DetectionObject	CodeCode Available	1
OK-VQA: A Visual Question Answering Benchmark Requiring External Knowledge	May 31, 2019	object-detectionObject Detection	CodeCode Available	1
GFF: Gated Fully Fusion for Semantic Segmentation	Apr 3, 2019	Scene ParsingScene Understanding	CodeCode Available	1
Curriculum Model Adaptation with Synthetic and Real Data for Semantic Foggy Scene Understanding	Jan 5, 2019	Domain AdaptationScene Understanding	CodeCode Available	1
Unified Perceptual Parsing for Scene Understanding	Jul 26, 2018	2D Semantic SegmentationScene Understanding	CodeCode Available	1
Visual Graphs from Motion (VGfM): Scene understanding with object geometry reasoning	Jul 16, 2018	3d scene graph generationGraph Generation	CodeCode Available	1
Digging Into Self-Supervised Monocular Depth Estimation	Jun 4, 2018	Camera Pose EstimationDepth Estimation	CodeCode Available	1
DeepScores -- A Dataset for Segmentation, Detection and Classification of Tiny Objects	Mar 27, 2018	General ClassificationObject	CodeCode Available	1
Semantic Line Detection and Its Applications	Oct 1, 2017	ClassificationGeneral Classification	CodeCode Available	1
LinkNet: Exploiting Encoder Representations for Efficient Semantic Segmentation	Jun 14, 2017	GPUScene Understanding	CodeCode Available	1
ScanNet: Richly-annotated 3D Reconstructions of Indoor Scenes	Feb 14, 2017	3D Object ClassificationGeneral Classification	CodeCode Available	1
Joint 2D-3D-Semantic Data for Indoor Scene Understanding	Feb 3, 2017	Scene Understanding	CodeCode Available	1

Show:10 25 50

← PrevPage 19 of 69Next →

All datasets Semantic Scene Understanding Challenge (passive actuation & ground-truth localisation)ADE20K val Semantic Scene Understanding Challenge (active actuation & ground-truth localisation)

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	ACRV Baseline	OMQ	0.44	—	Unverified
2	Team VGAI (TCS Research)	OMQ	0.37	—	Unverified
3	Demo_semantic_SLAM	OMQ	0.11	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	CPN(ResNet-101)	Mean IoU	46.3	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	ACRV Baseline	OMQ	0.35	—	Unverified