Architecture search of dynamic cells for semantic video segmentation

Nekrasov, V.; Chen, H.; Shen, C.; Reid, I.D.

Please use this identifier to cite or link to this item: https://hdl.handle.net/2440/129920

Scopus	Web of Science®	Altmetric
Citations
?	?

Full metadata record

DC Field	Value	Language
dc.contributor.author	Nekrasov, V.	-
dc.contributor.author	Chen, H.	-
dc.contributor.author	Shen, C.	-
dc.contributor.author	Reid, I.D.	-
dc.date.issued	2020	-
dc.identifier.citation	Proceedings of the IEEE Winter Conference on Applications of Computer Vision (WACV '20), 2020, pp.1959-1968	-
dc.identifier.isbn	9781728165530	-
dc.identifier.issn	2472-6737	-
dc.identifier.issn	2642-9381	-
dc.identifier.uri	http://hdl.handle.net/2440/129920	-
dc.description.abstract	In semantic video segmentation the goal is to acquire consistent dense semantic labelling across image frames. To this end, recent approaches have been reliant on manually arranged operations applied on top of static semantic segmentation networks – with the most prominent building block being the optical flow able to provide information about scene dynamics. Related to that is the line of research concerned with speeding up static networks by approximating expensive parts of them with cheaper alternatives, while propagating information from previous frames. In this work we attempt to come up with generalisation of those methods, and instead of manually designing contextual blocks that connect per-frame outputs, we propose a neural architecture search solution, where the choice of operations together with their sequential arrangement are being predicted by a separate neural network. We showcase that such generalisation leads to stable and accurate results across common benchmarks, such as CityScapes and CamVid datasets. Importantly, the proposed methodology takes only 2 GPU-days, finds high-performing cells and does not rely on the expensive optical flow computation.	-
dc.description.statementofresponsibility	Vladimir Nekrasov, Hao Chen, Chunhua Shen, Ian Reid	-
dc.language.iso	en	-
dc.publisher	IEEE	-
dc.relation.ispartofseries	IEEE Winter Conference on Applications of Computer Vision	-
dc.rights	©2020 IEEE	-
dc.source.uri	https://ieeexplore.ieee.org/xpl/conhome/9087828/proceeding	-
dc.title	Architecture search of dynamic cells for semantic video segmentation	-
dc.type	Conference paper	-
dc.contributor.conference	IEEE Winter Conference on Applications of Computer Vision (WACV) (1 Mar 2020 - 5 Mar 2020 : Snowmass Village, CO, USA)	-
dc.identifier.doi	10.1109/WACV45572.2020.9093531	-
dc.publisher.place	online	-
dc.relation.grant	http://purl.org/au-research/grants/arc/CE140100016	-
pubs.publication-status	Published	-
dc.identifier.orcid	Nekrasov, V. [0000-0001-9653-7539]	-
dc.identifier.orcid	Reid, I.D. [0000-0001-7790-6423]	-
Appears in Collections:	Aurora harvest 8 Computer Science publications

Files in This Item:

There are no files associated with this item.

Show simple item record

Adelaide Research & Scholarship