Publication: Video frame prediction via deep learning
| dc.conference.date | OCT 05-07, 2020 | |
| dc.conference.location | ELECTR NETWORK | |
| dc.conference.organizer | 2020 28th Signal Processing and Communications Applications Conference (SIU) | |
| dc.contributor.department | Department of Electrical and Electronics Engineering | |
| dc.contributor.facultymember | Yes | |
| dc.contributor.kuauthor | Tekalp, Ahmet Murat | |
| dc.contributor.kuauthor | Yılmaz, Mustafa Akın | |
| dc.contributor.schoolcollegeinstitute | College of Engineering | |
| dc.date.accessioned | 2024-11-09T23:13:11Z | |
| dc.date.issued | 2020 | |
| dc.description.abstract | This paper provides new results over our previous work presented in ICIP 2019 on the performance of learned frame prediction architectures and associated training methods. More specifically, we show that using an end-to-end residual connection in the fully convolutional neural network (FCNN) provides improved performance. in order to provide comparative results, we trained a residual FCNN, A convolutional RNN (CRNN), and a convolutional long-short term memory (CLSTM) network for next frame prediction using the mean square loss. We performed both stateless and stateful training for recurrent networks. Experimental results show that the residual FCNN architecture performs the best in terms of peak signal to noise ratio (PSNR) at the expense of higher training and test (inference) computational complexity. the CRNN can be stably and efficiently trained using the stateful truncated backpropagation through time procedure, and requires an order of magnitude less inference runtime to achieve an acceptable performance in near real-time. | |
| dc.description.fulltext | No | |
| dc.description.harvestedfrom | Manual | |
| dc.description.indexedby | WOS | |
| dc.description.openaccess | NO | |
| dc.description.peerreviewstatus | N/A | |
| dc.description.publisherscope | International | |
| dc.description.readpublish | N/A | |
| dc.description.sponsoredbyTubitakEu | TÜBİTAK | |
| dc.description.sponsorship | This work was supported by TÜBİTAK under project no. 217E033. A. Murat Tekalp also acknowledges support from the Turkish Academy of Sciences (TÜBA). | |
| dc.description.studentonlypublication | No | |
| dc.description.studentpublication | Yes | |
| dc.description.version | N/A | |
| dc.identifier.WoSQuartile | N/A | |
| dc.identifier.embargo | N/A | |
| dc.identifier.grantno | 217E033 | |
| dc.identifier.isbn | 9781728172064 | |
| dc.identifier.issn | 2165-0608 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.14288/9946 | |
| dc.identifier.wos | 000653136100021 | |
| dc.keywords | Frame prediction | |
| dc.keywords | Deep learning | |
| dc.keywords | Recurrent network architectures | |
| dc.keywords | Stateful training | |
| dc.keywords | Convolutional network architectures | |
| dc.language.iso | tur | |
| dc.publisher | Institute of Electrical and Electronics Engineers | |
| dc.relation.affiliation | Koç University | |
| dc.relation.collection | Koç University Institutional Repository | |
| dc.relation.ispartof | Signal Processing and Communications Applications Conference | |
| dc.relation.openaccess | N/A | |
| dc.rights | N/A | |
| dc.subject | Civil engineering | |
| dc.subject | Electrical electronics engineering | |
| dc.subject | Telecommunication | |
| dc.title | Video frame prediction via deep learning | |
| dc.type | Conference Proceeding | |
| dspace.entity.type | Publication | |
| local.contributor.kuauthor | Yılmaz, Mustafa Akın | |
| local.contributor.kuauthor | Tekalp, Ahmet Murat | |
| relation.isOrgUnitOfPublication | 21598063-a7c5-420d-91ba-0cc9b2db0ea0 | |
| relation.isOrgUnitOfPublication.latestForDiscovery | 21598063-a7c5-420d-91ba-0cc9b2db0ea0 | |
| relation.isParentOrgUnitOfPublication | 8e756b23-2d4a-4ce8-b1b3-62c794a8c164 | |
| relation.isParentOrgUnitOfPublication.latestForDiscovery | 8e756b23-2d4a-4ce8-b1b3-62c794a8c164 |
