Publication: A computationally efficient piecewise linear training algorithm for neural networks utilizing continuous special ordered sets
| dc.contributor.department | Department of Industrial Engineering | |
| dc.contributor.department | Department of Chemical and Biological Engineering | |
| dc.contributor.department | Graduate School of Sciences and Engineering | |
| dc.contributor.kuauthor | Köksal, Ece Serenat | |
| dc.contributor.kuauthor | Türkay, Metin | |
| dc.contributor.kuauthor | Aydın, Erdal | |
| dc.contributor.schoolcollegeinstitute | GRADUATE SCHOOL OF SCIENCES AND ENGINEERING | |
| dc.contributor.schoolcollegeinstitute | College of Engineering | |
| dc.date.accessioned | 2026-07-17T08:28:30Z | |
| dc.date.issued | 2026 | |
| dc.description.abstract | Artificial neural networks are commonly employed for data-driven modelling of complex nonlinear processes; however, their training may be hindered by the nonlinearity of activation functions and the dependence on local solvers. Obtaining an efficient global solution for neural network training continuous to be an unresolved challenge. A common strategy involves approximating activation functions using piecewise linear formulations to convexify the problem, however this often results in high computational costs due to the addition of auxiliary binary variables. This study proposes integrating piecewise linear formulations for neural network activation functions with a tailored branching algorithm that explores efficient linear programming relaxations to effectively explore the solution space. In the proposed framework, a subset of network parameters is obtained via regular, gradient-descent based training and kept fixed during the training process, while the remaining parameters are determined through the proposed formulation. Unlike conventional mixed-integer programming approaches, where the reliance on binary variables increases the computational complexity, the applied continuous special ordered set method achieves polynomial growth in computation, thereby ensuring better scalability for larger problem instances. The proposed method achieves minimal training error while significantly reducing CPU time. Experiments on datasets of equal dimensionality confirm that the efficiency of the algorithm is not dataset-specific, demonstrating consistent CPU time trends across multiple datasets. These findings highlight the generalizability of the suggested method and its potential to enhance artificial neural network training efficiency across various applications. | |
| dc.description.harvestedfrom | Manual | |
| dc.description.indexedby | WOS | |
| dc.description.indexedby | Scopus | |
| dc.description.publisherscope | International | |
| dc.description.readpublish | N/A | |
| dc.description.sponsoredbyTubitakEu | N/A | |
| dc.description.version | Published Version | |
| dc.identifier.ScopusPercentile | 79 | |
| dc.identifier.ScopusQuartile | Q1 | |
| dc.identifier.WoSPercentile | 58.2 | |
| dc.identifier.WoSQuartile | Q2 | |
| dc.identifier.doi | 10.1016/j.cherd.2026.03.047 | |
| dc.identifier.eissn | 1744-3563 | |
| dc.identifier.embargo | N/A | |
| dc.identifier.endpage | 277 | |
| dc.identifier.issn | 0263-8762 | |
| dc.identifier.scopus | 2-s2.0-105035242575 | |
| dc.identifier.startpage | 266 | |
| dc.identifier.uri | http://doi.org/10.1016/j.cherd.2026.03.047 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.14288/33378 | |
| dc.identifier.volume | 229 | |
| dc.identifier.wos | 001742690100002 | |
| dc.keywords | Artificial neural networks | |
| dc.keywords | Mixed integer linear programming | |
| dc.keywords | Piecewise linear functions | |
| dc.keywords | Special ordered set variables | |
| dc.keywords | Quasi- convex formulation | |
| dc.language | eng | |
| dc.publisher | Elsevier | |
| dc.relation.affiliation | Koç University | |
| dc.relation.collection | Koç University Institutional Repository | |
| dc.relation.ispartof | Chemical Engineering Research and Design | |
| dc.relation.openaccess | N/A | |
| dc.rights | N/A | |
| dc.rights.uri | N/A | |
| dc.subject | Engineering | |
| dc.subject | Chemical | |
| dc.title | A computationally efficient piecewise linear training algorithm for neural networks utilizing continuous special ordered sets | |
| dc.type | Journal Article | |
| dspace.entity.type | Publication | |
| relation.isOrgUnitOfPublication | d6d00f52-d22d-4653-99e7-863efcd47b4a | |
| relation.isOrgUnitOfPublication | c747a256-6e0c-4969-b1bf-3b9f2f674289 | |
| relation.isOrgUnitOfPublication | 3fc31c89-e803-4eb1-af6b-6258bc42c3d8 | |
| relation.isOrgUnitOfPublication.latestForDiscovery | d6d00f52-d22d-4653-99e7-863efcd47b4a | |
| relation.isParentOrgUnitOfPublication | 434c9663-2b11-4e66-9399-c863e2ebae43 | |
| relation.isParentOrgUnitOfPublication | 8e756b23-2d4a-4ce8-b1b3-62c794a8c164 | |
| relation.isParentOrgUnitOfPublication.latestForDiscovery | 434c9663-2b11-4e66-9399-c863e2ebae43 |
