Computer vision
39 milestones used this technique.
Waymo One Launches Fully Driverless Rides to the General Public in Phoenix, Arizona
On 8 October 2020, Waymo opened its Waymo One ride-hailing service to the general public in the greater Phoenix, Arizona area, operating without a safety driver in the vehicle, the first time a commercial autonomous vehicle service had done so at public scale.
Facebook AI Research Publishes GrokNet, a Unified Computer Vision Model for Commerce Understanding
In 2020, researchers at Facebook AI published GrokNet, a unified deep learning system for product understanding in commerce settings, capable of recognising object categories, attributes such as colour and material, and brand information from product images at scale across Facebook Shops.
Facebook AI Research Releases Detectron2
In October 2019, Facebook AI Research released Detectron2, an open-source object detection and segmentation framework built on PyTorch, supporting algorithms including Mask R-CNN, DensePose, and panoptic feature pyramid networks, replacing the earlier Caffe2-based Detectron.
Textron Systems Unveils Ripsaw M5 Robotic Combat Vehicle
In October 2019, Textron Systems unveiled the Ripsaw M5 Robotic Combat Vehicle at the Association of the United States Army Annual Meeting in Washington, D.C., demonstrating an unmanned ground vehicle with autonomous navigation, 360-degree situational awareness, and configurable mission payloads.
LOVOT Companion Robot Unveiled by Groove X
In December 2018, Groove X, a Japanese robotics company founded by Kaname Hayashi, unveiled LOVOT, a companion robot designed to elicit emotional attachment rather than perform practical tasks, equipped with more than 50 sensors, a thermal camera array, and a neural-processing unit to recognise and respond to human behaviour.
Waymo One Commercial Ride-Hailing Service Launch
In December 2018, Waymo LLC launched Waymo One, a fare-charging autonomous ride-hailing service operating in the Greater Phoenix, Arizona area, marking the first time a driverless vehicle service had been made available to paying members of the public in the United States.
CIMON Launched to the International Space Station
In June 2018, DLR, Airbus, and IBM launched CIMON (Crew Interactive Mobile Companion), a spherical, voice-controlled AI assistant, to the International Space Station aboard SpaceX CRS-15. It was designed to support ESA astronaut Alexander Gerst with procedural tasks and reduce cognitive workload.
Movidius (Intel) Launches Neural Compute Stick
In July 2017, Movidius, an Intel subsidiary, released the Movidius Neural Compute Stick, a USB-form-factor device housing the Myriad 2 Vision Processing Unit, enabling developers to run inference from trained deep neural networks on low-power edge hardware without a remote server.
Facebook Deploys AI-Based Photo and Video Integrity Systems to Detect Nudity and Graphic Violence at Scale
From at least 2016, Facebook applied convolutional neural network-based computer vision systems to automatically detect nudity and graphic violence across photos and videos uploaded to its platform, processing billions of pieces of content as part of its scaled content integrity infrastructure.
Moley Robotics Demonstrates Robotic Kitchen System at Hannover Messe
In April 2015, London-based Moley Robotics unveiled a prototype robotic kitchen system at Hannover Messe, comprising a pair of dexterous robotic arms capable of replicating recorded human cooking movements, integrated with an oven, hob, and dishwasher in a fitted kitchen unit.
SoftBank Robotics and Aldebaran Unveil Pepper, a Humanoid Robot with Emotion Recognition
In June 2014, SoftBank Robotics and its subsidiary Aldebaran Robotics unveiled Pepper, a 1.2-metre humanoid robot equipped with an emotion-recognition system capable of detecting human facial expressions, voice tone, and body language, intended for retail and customer-service deployment.
Xinlei Chen, Abhinav Shrivastava and Abhinav Gupta at Carnegie Mellon University present NEIL (Never-Ending Image Learner) at ICCV 2013
In December 2013, Xinlei Chen, Abhinav Shrivastava and Abhinav Gupta at Carnegie Mellon University presented NEIL (Never-Ending Image Learner) at ICCV 2013, a continuously running system that autonomously mined semantic relationships between visual concepts from unlabelled web images without human supervision.
AeroVironment Nano Hummingbird: DARPA-funded Flapping-Wing Micro Air Vehicle
In February 2011, AeroVironment unveiled the Nano Hummingbird, a DARPA-funded flapping-wing micro air vehicle weighing 19 grams, less than a AA battery, that carried a video camera and used control systems to mimic hummingbird flight, including hover and omnidirectional movement.
Microsoft Research Develops Real-Time Human Pose Estimation for Kinect
In 2011, Jamie Shotton and colleagues at Microsoft Research Cambridge published a method for real-time human pose estimation from a single depth image, using randomised decision forests trained on synthetic data. The technique powered the skeleton-tracking feature of Microsoft Kinect and was presented at CVPR 2011.
Autonomous Robotic Cardiac Surgery Guided by Machine Learning
In May 2006, a robotic surgical system at the University of Toronto, trained on data from more than 10,000 prior operations, performed an autonomous 50-minute cardiac procedure on a beating human heart, demonstrating machine-learning-guided autonomy in a clinical surgical setting.
Dasarobot Genibo QD Consumer Pet Robot
In April 2006, South Korean company Dasarobot unveiled the Genibo QD, a consumer pet robot modelled on a dog and equipped with cameras, infrared sensors, and voice-recognition software designed to simulate emotional responses to its owner.
Stanford Racing Team's Stanley Wins DARPA Grand Challenge 2005
On 8 October 2005, a Stanford University team led by Sebastian Thrun entered Stanley, a modified Volkswagen Touareg, in the DARPA Grand Challenge. Stanley completed the 131.6-mile (211.8 km) Mojave Desert course autonomously in under 7 hours, finishing first and winning the $2 million prize.
NASA Mars Exploration Rovers Spirit and Opportunity Begin Autonomous Surface Operations
NASA's Mars Exploration Rovers Spirit and Opportunity landed on Mars in January 2004 and used onboard autonomous navigation software to traverse the Martian surface, making decisions about safe paths without real-time human control due to communication delays of up to 20 minutes.
Rapid Object Detection Using a Boosted Cascade of Simple Features (Viola–Jones Face Detection)
Paul Viola and Michael Jones, then at Compaq CRL and Mitsubishi Electric Research Laboratories respectively, published a cascaded boosting framework for real-time face detection, first presented at CVPR in December 2001 and consolidated in the International Journal of Computer Vision in 2004. The method ran at frame rates suitable for live video on consumer hardware.
TEXTAL System for AI-Assisted Automated Protein Model Building
In 2003, Thomas R. Ioerger and James C. Sacchettini at Texas A&M University described TEXTAL, a pattern-recognition system that automatically traced atomic models through crystallographic electron density maps, substantially reducing the manual labour required in protein structure determination.
Bag of Words Applied to Computer Vision (Visual Vocabulary / Bag of Visual Words)
Josef Sivic and Andrew Zisserman at the University of Oxford applied the Bag of Words text-retrieval model to visual features in their 2003 ICCV paper 'Video Google', representing image regions as a vocabulary of visual words to enable efficient object retrieval from video.
Autographer: Autonomous Robot Photographer Deployed at AAAI/IAAI 2003
In 2003, Selene Mota and Rosalind Picard at the MIT Media Lab deployed an autonomous mobile robot called Autographer at IAAI 2003. Over five days it navigated a conference environment, interacted with roughly 5,000 people, and captured more than 3,000 photographs, 35% of which recipients requested by email.
Structured Light for Robust Correspondence in Active Stereo Vision
In 2002, Li Zhang, Brian Curless, and Steven M. Seitz at the University of Washington presented a method using structured light patterns projected onto scenes to establish robust stereo correspondences, enabling reliable 3D reconstruction under conditions where passive stereo fails.
FDA Clears CyberKnife for Full-Body Tumour Treatment
In August 2001, the US Food and Drug Administration cleared the CyberKnife Robotic Radiosurgery System for treatment of tumours anywhere in the body, extending an earlier 1999 clearance limited to the head and neck. The system uses real-time image guidance and robotic positioning to deliver radiation with sub-millimetre accuracy.
Space Station Remote Manipulator System (Canadarm2) Deployed on the International Space Station
On 22 April 2001, the Space Station Remote Manipulator System (Canadarm2), built by MD Robotics of Canada, was installed on the International Space Station during the STS-100 mission. The 17-metre robotic arm used computer-vision and control-systems software to manoeuvre equipment, modules, and astronauts autonomously or under remote supervision.
Hawk-Eye Ball-Tracking System Developed by Paul Hawkins at Roke Manor Research
In 2001, Paul Hawkins, working at Roke Manor Research in Hampshire, developed Hawk-Eye, a computer-vision system that triangulates footage from multiple broadcast cameras to reconstruct the three-dimensional trajectory of a sports ball in near real time, first used in cricket television coverage.
Kismet: Sociable Robot Developed by Cynthia Breazeal at MIT
Around 2000, Cynthia Breazeal at the MIT Artificial Intelligence Laboratory completed Kismet, a robotic head capable of perceiving and expressing emotion through coordinated facial features, pioneering research into socially intelligent robots able to engage in natural affective interaction with humans.
Intelligent Room and Affective Computing Agents at MIT AI Laboratory
In 1998, researchers at the MIT Artificial Intelligence Laboratory, including Rodney Brooks and Cynthia Breazeal, developed the Intelligent Room project alongside work on emotionally expressive robotic agents, combining layered behaviour-based architectures with computer vision to enable a physical space to perceive and respond to human occupants.
Sojourner Rover Lands on Mars as First Autonomous Wheeled Vehicle on Another Planet
On 4 July 1997, NASA's Sojourner rover became the first wheeled vehicle to operate on another planet, driving onto the Martian surface as part of the Mars Pathfinder mission. Sojourner used onboard hazard-avoidance logic and laser stripe sensors to navigate autonomously when out of direct communication with Earth.
Ernst Dickmanns and VaMoRs Autonomous Van Test, Bundeswehr University Munich
In 1986, Ernst Dickmanns and colleagues at Bundeswehr University Munich demonstrated VaMoRs, a Mercedes-Benz van retrofitted with cameras and real-time computer vision, capable of autonomous driving on traffic-free roads at speeds up to approximately 96 km/h, marking one of the earliest working autonomous vehicle demonstrations.
WABOT-2 Humanoid Robot Demonstrated at Waseda University
In 1984, researchers at Waseda University in Japan completed WABOT-2, a humanoid robot capable of reading printed musical scores, communicating with a human performer via speech, and playing an electronic organ using its fingers and foot pedals at a level comparable to an average adult pianist.
Primal Sketch Theory of Early Visual Representation Described by David Marr
David Marr, working at MIT's Artificial Intelligence Laboratory, formalised the primal sketch as the first stage of his three-level theory of visual processing, published posthumously in 'Vision' (1982). The model proposed that the visual system constructs a symbolic, viewer-centred description of intensity changes and local geometry before any object recognition takes place.
Stanford Cart Slider Modification by Hans Moravec
During the late 1970s, Hans Moravec at Stanford University extended the Stanford Cart, originally built by James L. Adams in the 1960s, with a sliding camera mount, enabling the robot to navigate autonomously across a chair-filled room in 1979 using stereo vision, a milestone in autonomous vehicle research.
FREDDY II Robot, University of Edinburgh Department of Machine Intelligence and Perception
By 1973, researchers at the University of Edinburgh's Department of Machine Intelligence and Perception had developed FREDDY II, a robot arm system that used computer vision and tactile feedback to identify and assemble simple objects from a pile of scattered parts, demonstrating integrated perception and manipulation in a single robotic system.
Stanford Cart Road-Following Experiment Under Les Earnest
In late 1971, Les Earnest at the Stanford Artificial Intelligence Laboratory adapted the Stanford Cart for autonomous road-following, using a television camera and a low-power radio control link to guide the vehicle along a road at approximately 0.8 mph (1.3 kph).
SRI International Begins Development of Shakey the Robot
From 1966, researchers at the Stanford Research Institute (led by Charles Rosen and including Nils Nilsson, Bertram Raphael, and Peter Hart) developed Shakey, a mobile robot that combined computer vision, natural language input, and automated planning to navigate and manipulate objects in a real environment.
Machine Perception of Three-Dimensional Solids, Lawrence Gilman Roberts (MIT Lincoln Laboratory)
In 1963, Lawrence Gilman Roberts, working at MIT Lincoln Laboratory, completed his doctoral thesis demonstrating that a computer could interpret a 2D photograph of polyhedral objects, reconstruct their 3D structure, and re-render them from arbitrary viewpoints with hidden lines removed, establishing foundational methods for machine interpretation of three-dimensional scenes.
Stanford Cart Radio-Link Configuration (1963)
In September 1963, researchers at Stanford University fitted the Stanford Cart with an analogue computer and radio-control links, allowing a remote operator to steer the vehicle using a television camera feed and a displayed target dot, in an early attempt at closed-loop vision-guided vehicle control.
Stanford Cart (Cable-Controlled Version)
In 1961, James L. Adams at Stanford University built the first version of the Stanford Cart, a four-wheeled vehicle tethered by cable to a remote console and television monitor, to investigate video-guided remote control. Tests showed the Cart could not operate reliably above approximately 0.2 mph owing to communication delays introduced by the cable link.