Probabilistic Future Prediction for Video Scene Understanding
cam.orpheus.counter | 29 | * |
dc.contributor.author | Hu, Anthony | |
dc.contributor.author | Cotter, Fergal | |
dc.contributor.author | Mohan, Nikhil | |
dc.contributor.author | Gurau, Corina | |
dc.contributor.author | Kendall, Alex | |
dc.date.accessioned | 2021-10-20T23:30:09Z | |
dc.date.available | 2021-10-20T23:30:09Z | |
dc.description.abstract | We present a novel deep learning architecture for probabilistic future prediction from video. We predict the future semantics, geometry and motion of complex real-world urban scenes and use this representation to control an autonomous vehicle. This work is the first to jointly predict ego-motion, static scene, and the motion of dynamic agents in a probabilistic manner, which allows sampling consistent, highly probable futures from a compact latent space. Our model learns a representation from RGB video with a spatio-temporal convolutional module. The learned representation can be explicitly decoded to future semantic segmentation, depth, and optical flow, in addition to being an input to a learnt driving policy. To model the stochasticity of the future, we introduce a conditional variational approach which minimises the divergence between the present distribution (what could happen given what we have seen) and the future distribution (what we observe actually happens). During inference, diverse futures are generated by sampling from the present distribution. | |
dc.description.sponsorship | Toshiba Europe, grant G100453 | |
dc.identifier.doi | 10.17863/CAM.77127 | |
dc.identifier.uri | https://www.repository.cam.ac.uk/handle/1810/329681 | |
dc.language.iso | eng | |
dc.rights | All rights reserved | |
dc.rights.uri | http://www.rioxx.net/licenses/all-rights-reserved | |
dc.title | Probabilistic Future Prediction for Video Scene Understanding | |
dc.type | Conference Object | |
dcterms.dateAccepted | 2020-07-03 | |
pubs.conference-finish-date | 2020-08-28 | |
pubs.conference-name | European Conference on Computer Vision (ECCV) | |
pubs.conference-start-date | 2020-08-23 | |
rioxxterms.licenseref.startdate | 2020-07-03 | |
rioxxterms.licenseref.uri | http://www.rioxx.net/licenses/all-rights-reserved | |
rioxxterms.type | Conference Paper/Proceeding/Abstract | |
rioxxterms.version | AM | |
rioxxterms.versionofrecord | 10.17863/CAM.77127 |
Files
Original bundle
1 - 1 of 1
Loading...
- Name:
- ECCV2020___Probabilistic_Future_Prediction_for_Video_Scene_Understanding.pdf
- Size:
- 4.94 MB
- Format:
- Adobe Portable Document Format
- Description:
- Accepted version
- Licence
- http://www.rioxx.net/licenses/all-rights-reserved
License bundle
1 - 1 of 1
No Thumbnail Available
- Name:
- DepositLicenceAgreementv2.1.pdf
- Size:
- 150.9 KB
- Format:
- Adobe Portable Document Format