Publication: Robust lip-motion features for speaker identification
| dc.conference.date | MAR 19-23, 2005 | |
| dc.conference.location | Philadelphia, PA | |
| dc.conference.organizer | Institute of Electrical and Electronics Engineers | |
| dc.contributor.department | MVGL (Multimedia, Vision and Graphics Laboratory) | |
| dc.contributor.facultymember | Yes | |
| dc.contributor.kuauthor | Çetingül, Hasan Ertan | |
| dc.contributor.kuauthor | Erzin, Engin | |
| dc.contributor.kuauthor | Yemez, Yücel | |
| dc.contributor.kuauthor | Tekalp, Ahmet Murat | |
| dc.contributor.schoolcollegeinstitute | Laboratory | |
| dc.date.accessioned | 2024-11-10T00:06:07Z | |
| dc.date.issued | 2005 | |
| dc.description.abstract | This paper addresses the selection of robust lip-motion features for audio-visual open-set speaker identification problem. We consider two alternatives for initial lip motion representation. In the first alternative. the feature vector is composed of the 2D-DCT coefficients of the motion vectors estimated within the detected rectangular mouth region whereas in the second, lip boundaries are tracked over the video frames and only the motion vectors around the lip contour are taken into account along with the shape of the lip boundary. Experimental results of the HMM-based identification system are included for performance comparison of the two lip motion representation alternatives. | |
| dc.description.fulltext | Yes | |
| dc.description.harvestedfrom | Manual | |
| dc.description.indexedby | WOS | |
| dc.description.indexedby | Scopus | |
| dc.description.openaccess | Green OA | |
| dc.description.peerreviewstatus | N/A | |
| dc.description.publisherscope | International | |
| dc.description.readpublish | N/A | |
| dc.description.sponsoredbyTubitakEu | N/A | |
| dc.description.studentonlypublication | No | |
| dc.description.studentpublication | Yes | |
| dc.description.version | Post-print | |
| dc.identifier.WoSQuartile | N/A | |
| dc.identifier.embargo | No | |
| dc.identifier.filenameinventoryno | IR06898 | |
| dc.identifier.isbn | 0780388747 | |
| dc.identifier.issn | 1520-6149 | |
| dc.identifier.scopus | 2-s2.0-33646818965 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.14288/16559 | |
| dc.identifier.wos | 000229404200128 | |
| dc.keywords | Speech | |
| dc.keywords | Lip tracking | |
| dc.keywords | Machine vision | |
| dc.language.iso | eng | |
| dc.publisher | IEEE | |
| dc.relation.affiliation | Koç University | |
| dc.relation.collection | Koç University Institutional Repository | |
| dc.relation.ispartof | 2005 IEEE International Conference On Acoustics, Speech, And Signal Processing, Vols 1-5: Speech Processing | |
| dc.relation.openaccess | Yes | |
| dc.rights | Other | |
| dc.subject | Computer science | |
| dc.subject | Artificial intelligence | |
| dc.subject | Engineering | |
| dc.subject | Electrical electronic engineering | |
| dc.title | Robust lip-motion features for speaker identification | |
| dc.type | Conference Proceeding | |
| dspace.entity.type | Publication | |
| local.contributor.kuauthor | Çetingül, Hasan Ertan | |
| local.contributor.kuauthor | Yemez, Yücel | |
| local.contributor.kuauthor | Erzin, Engin | |
| local.contributor.kuauthor | Ahmet Murat | |
| relation.isOrgUnitOfPublication | cb6bbbf6-fd19-4052-b581-f591a9748d21 | |
| relation.isOrgUnitOfPublication.latestForDiscovery | cb6bbbf6-fd19-4052-b581-f591a9748d21 | |
| relation.isParentOrgUnitOfPublication | 20385dee-35e7-484b-8da6-ddcc08271d96 | |
| relation.isParentOrgUnitOfPublication.latestForDiscovery | 20385dee-35e7-484b-8da6-ddcc08271d96 |
Files
Original bundle
1 - 1 of 1
