Microsoft Coco
Common Objects in Context
Bibliographic Data
| ID | 23324941 |
|---|---|
| Authors | Tsung-Yi Lin (0000-0003-4819-0627, Cornell University), Michael Maire (0000-0002-9778-6673, California Institute of Technology), Serge Belongie (0000-0002-0388-5217, Cornell University), James Hays (0000-0001-7016-4252), Pietro Perona (0000-0002-7583-5809, California Institute of Technology), Deva Ramanan (0009-0008-9180-8983, UC Irvine Health), Piotr Dollár (Microsoft (United States)), C Lawrence Zitnick (0009-0006-2032-9374, Microsoft (United States)) |
| Year | 2014 |
| Pages | 740-755 |
| Publication date | 2014-01-01 |
| Peer Reviewed | Yes |
| Open Access | Yes |
| Type | CHAPTER |
| Venue | Computer Vision -- ECCV 2014 (SOURCE_BOOK) |
| Publisher | Springer International Publishing (PUBLISHER • SG) |
| DOI | 10.1007/978-3-319-10602-1_48 |
| OpenAlex | W1861492603 |
| ISBN | 9783319106021 |
| Language | EN |
| Citations received | 149 |
| References cited | 35 |
Baseline (sea) · Bounding overwatch · Computer vision · Context (archaeology) · Geography · Image (mathematics) · Image segmentation · Minimum bounding box · Object (grammar) · Object detection · Pascal (unit) · Pattern recognition (psychology) · Segmentation · Spotting · Advanced Image and Video Retrieval Techniques · Advanced Neural Network Applications · Artificial Intelligence · Computer Science · Multimodal Machine Learning Applications
Enhancing Social Agency and Inclusion for Severe Speech and Motor Impaired Individuals through Gamified, Eye-Gaze Controlled Socially Assistive Robotics
Deep Learning for Detection of Underwater Aircraft Wrecks from US Conflicts
Combined Detection and Segmentation of Archeological Structures from LiDAR Data Using a Deep Learning Approach
Objektintunnistus ja suomalainen elokuva
Deep Representation Learning With Full Center Loss for Credit Card Fraud Detection
Design and Development of an Internet of Smart Cameras Solution for Complex Event Detection in Covid-19 Risk Behaviour Recognition
Mood Themes the World
Image-To-Insight
GeoSight
A framework to enhance disaster debris estimation with AI and aerial photogrammetry
Social distance monitoring framework using deep learning architecture to control infection transmission of Covid-19 pandemic
Pedestrian tracking in outdoor spaces of a suburban university campus for the investigation of occupancy patterns
Computational Photography
Reading Assistant for Visually Challenged Peoples with Advance Image Capturing Technique Using Machine Learning
A Work-Related Musculoskeletal Disorders (WMSDs) Risk-Assessment System Using a Single-View Pose Estimation Model
Deep Learning-Based Object Detection, Localisation and Tracking for Smart Wheelchair Healthcare Mobility
Deep learning to detect built cultural heritage from satellite imagery. - Spatial distribution and size of vernacular houses in Sumba, Indonesia
DEArt
Deep learning-based automated tile defect detection system for Portuguese cultural heritage buildings
Out of dataset, out of algorithm, out of mind
Representation of Linguistic Form and Function in Recurrent Neural Networks
Eye in the sky
Understanding cities with machine eyes
Detecting older pedestrians and aging-friendly walkability using computer vision technology and street view imagery
GeoAI in terrain analysis
Flood depth mapping in street photos with image processing and deep neural networks
Semantic Riverscapes
DeepLab
GOd, mOther and sOldier
Searching for Computer Vision North Stars
Scene context is predictive of unconstrained object similarity judgments
The Digital Humanities and the Ladino Press
Geo‐Foundation Models
Disconnecting with Social Networking Sites
Prostitution and sex work
High-Resolution Image Synthesis with Latent Diffusion Models
Social media data for conservation science
Squeeze-and-Excitation Networks
The Cityscapes Dataset for Semantic Urban Scene Understanding
Mask R-CNN
SegNet
Building instance classification using street view images
Focal Loss for Dense Object Detection
OpenPose
Machine Learning for Cultural Heritage
Faster R-CNN
Encoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentation
Interpreting Black-Box Models
Places
Artificial intelligence in the creative industries
ImageNet Large Scale Visual Recognition Challenge
Microsoft Coco
Machine behaviour
A Survey on Evaluation of Large Language Models
Semantic Understanding of Scenes Through the ADE20K Dataset
Global Research Trends of Artificial Intelligence on Histopathological Images
Automatic Movement Recognition for Evaluating the Gross Motor Development of Infants
Drone-Based Water Level Detection in Flood Disasters
Pollen-Yolo
Detection and Visualisation of Pneumoconiosis Using an Ensemble of Multi-Dimensional Deep Features Learned from Chest X-rays
Ensuring academic integrity through automated online exam proctoring a decade long systematic review
A Graph-Based Human-Centered GeoAI for Understanding Street-Level Environment and Traffic Accident Frequency with Street View Imagery
Visual storytelling through the void
Beijing’s urbanization reflected in apartment layouts
Distant viewing and multimodality theory
From Percepts to Semantics
Multilevel Ownership Protection via Watermarking and Encryption
ArtCap
Compressing the Multiobject Tracking Model via Knowledge Distillation
Twofold Structured Features-Based Siamese Network for Infrared Target Tracking
Social Equity and Inclusion With Fair Dataset Representation of Diverse Populations in Diffusion Models
SiamATTRPN
Flexible Dual-Branch Siamese Network
Zero-Shot Cross-Lingual Knowledge Transfer in VQA via Multimodal Distillation
Self-Prompt Guided Image Outpainting Model for Captions Absence in Social Scenes
Object Detector Based on Center Keypoints for Behavior Recognition in Classroom Scenes
A Novel Chaotic Map and Its Application to Secure Transmission of Multimodal Images
Deformable Blur Sensing and Regression Analysis ReID Feature Fusion for Multitarget Multicamera Tracking Systems in Highway Scenarios
Generative Adversarial Networks and Its Applications in Biomedical Informatics
Cloud and Snow Segmentation in Satellite Images Using an Encoder–Decoder Deep Convolutional Neural Networks
DeepWindows
A Research on Landslides Automatic Extraction Model Based on the Improved Mask R-CNN
Extracting the Urban Landscape Features of the Historic District from Street View Images Based on Deep Learning
Canopy Assessment of Cycling Routes
Building Block Extraction from Historical Maps Using Deep Object Attention Networks
EnvSLAM
Harnessing Foundation Models for Optical–SAR Object Detection via Gated–Guided Fusion
The Influence of Point Cloud Accuracy from Image Matching on Automatic Preparation of Training Datasets for Object Detection in UAV Images
Semantic Relation Model and Dataset for Remote Sensing Scene Understanding
A Lightweight Object Detection Method in Aerial Images Based on Dense Feature Fusion Path Aggregation Network
Big Data-Driven Pedestrian Analytics
Heri-Graphs
DeepDBSCAN
RDQS
Semantic Segmentation of Remote-Sensing Imagery Using Heterogeneous Big Data
RepDarkNet
Using Object Detection on Social Media Images for Urban Bicycle Infrastructure Planning
From Internet Meme to the Mainstream
Deep Learning for Visual Analytics of the Spread of Covid-19 Infection in Crowded Urban Environments
Validation of an Augmented Parcel Approach for Hurricane Regional Loss Assessments
| Unique citing works | 149 |
|---|---|
| Citations per year | 12,42 |
| Citation span | 2014 - 2026 (13) |
| Citation velocity | current |
| Highly cited | Yes |
| Citation types | Neutral: 118 |