Publication: Discriminative lip-motion features for biometric speaker identification
| dc.conference.date | OCT 24-27, 2004 | |
| dc.conference.location | Singapore, Singapore | |
| dc.conference.organizer | International Conference on Image Processing (ICIP 2004) | |
| dc.contributor.department | Department of Electrical and Electronics Engineering | |
| dc.contributor.department | Department of Computer Engineering | |
| dc.contributor.department | Graduate School of Sciences and Engineering | |
| dc.contributor.facultymember | Yes | |
| dc.contributor.kuauthor | Çetingül, Hasan Ertan | |
| dc.contributor.kuauthor | Erzin, Engin | |
| dc.contributor.kuauthor | Tekalp, Ahmet Murat | |
| dc.contributor.kuauthor | Yemez, Yücel | |
| dc.contributor.schoolcollegeinstitute | College of Engineering | |
| dc.contributor.schoolcollegeinstitute | GRADUATE SCHOOL OF SCIENCES AND ENGINEERING | |
| dc.date.accessioned | 2024-11-09T23:08:02Z | |
| dc.date.issued | 2004 | |
| dc.description.abstract | This paper addresses the selection of best lip motion features for biometric open-set speaker identification. The best features are those that result in the highest discrimination of individual speakers in a population. We first detect the face region in each video frame. The lip region for each frame is then segmented following registration of successive face regions by global motion compensation. The initial lip feature vector is composed of the 2D-DCT coefficients of the optical flow vectors within the lip region at each frame. The discriminant analysis is composed of two stages. At the first stage, the most discriminative features are selected from the full set of DCT coefficients of a single lip motion frame by using a probabilistic measure that maximizes the ratio of intra-class and inter-class probabilities. At the second stage, the resulting discriminative feature vectors are interpolated and concatenated for each time instant within a neighborhood, and further analyzed by LDA to reduce dimension, this time taking into account temporal discrimination information. Experimental results of the HMM-based speaker identification system are included to demonstrate the performance. | |
| dc.description.fulltext | No | |
| dc.description.harvestedfrom | Manual | |
| dc.description.indexedby | Scopus | |
| dc.description.indexedby | WOS | |
| dc.description.openaccess | YES | |
| dc.description.peerreviewstatus | N/A | |
| dc.description.publisherscope | International | |
| dc.description.readpublish | N/A | |
| dc.description.sponsoredbyTubitakEu | TÜBİTAK | |
| dc.description.sponsorship | This work ha7 been supported by TUBITAK under the project EEEAG-lOlE038 and by the European FP6 Network of Excellence SMILAR. | |
| dc.description.studentonlypublication | No | |
| dc.description.studentpublication | Yes | |
| dc.description.version | N/A | |
| dc.identifier.WoSQuartile | N/A | |
| dc.identifier.doi | 10.1109/ICIP.2004.1421480 | |
| dc.identifier.embargo | N/A | |
| dc.identifier.endpage | 2026 | |
| dc.identifier.grantno | TÜBİTAK [lOlE038] | |
| dc.identifier.isbn | 0780385543 | |
| dc.identifier.issn | 1522-4880 | |
| dc.identifier.scopus | 2-s2.0-20444432705 | |
| dc.identifier.startpage | 2023 | |
| dc.identifier.uri | https://doi.org/10.1109/ICIP.2004.1421480 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.14288/9250 | |
| dc.identifier.volume | 3 | |
| dc.identifier.wos | 000228043502150 | |
| dc.keywords | Biometry | |
| dc.keywords | Dimensions | |
| dc.keywords | Feature sets | |
| dc.keywords | Mean square errors | |
| dc.keywords | Database systems | |
| dc.keywords | Decision making | |
| dc.keywords | Eigenvalues and eigenfunctions | |
| dc.keywords | Errors | |
| dc.keywords | Interpolation | |
| dc.keywords | Mathematical models | |
| dc.keywords | Principal component analysis | |
| dc.keywords | Probability | |
| dc.keywords | Speech recognition | |
| dc.language.iso | eng | |
| dc.publisher | Institute of Electrical and Electronics Engineers (IEEE) | |
| dc.relation.affiliation | Koç University | |
| dc.relation.collection | Koç University Institutional Repository | |
| dc.relation.ispartof | Proceedings - International Conference on Image Processing, ICIP | |
| dc.relation.openaccess | N/A | |
| dc.rights | N/A | |
| dc.subject | Electrical electronics engineering | |
| dc.subject | Computer engineering | |
| dc.title | Discriminative lip-motion features for biometric speaker identification | |
| dc.type | Conference Proceeding | |
| dspace.entity.type | Publication | |
| local.contributor.kuauthor | Tekalp, Ahmet Murat | |
| local.contributor.kuauthor | Erzin, Engin | |
| local.contributor.kuauthor | Yemez, Yücel | |
| local.contributor.kuauthor | Çetingül, Hasan Ertan | |
| relation.isOrgUnitOfPublication | 21598063-a7c5-420d-91ba-0cc9b2db0ea0 | |
| relation.isOrgUnitOfPublication | 89352e43-bf09-4ef4-82f6-6f9d0174ebae | |
| relation.isOrgUnitOfPublication | 3fc31c89-e803-4eb1-af6b-6258bc42c3d8 | |
| relation.isOrgUnitOfPublication.latestForDiscovery | 21598063-a7c5-420d-91ba-0cc9b2db0ea0 | |
| relation.isParentOrgUnitOfPublication | 8e756b23-2d4a-4ce8-b1b3-62c794a8c164 | |
| relation.isParentOrgUnitOfPublication | 434c9663-2b11-4e66-9399-c863e2ebae43 | |
| relation.isParentOrgUnitOfPublication.latestForDiscovery | 8e756b23-2d4a-4ce8-b1b3-62c794a8c164 |
