Umeå University's logo

umu.sePublikasjoner
Endre søk
Link to record
Permanent link

Direct link
Prakhya, Karthik
Publikasjoner (3 av 3) Visa alla publikasjoner
Prakhya, K., Birdal, T. & Yurtsever, A. (2025). Convex formulations for training two-layer ReLU neural networks. In: 13th International Conference on Learning Representations, ICLR 2025: . Paper presented at International Conference on Learning Representations (ICLR), Singapore, April 24-28, 2025 (pp. 30682-30704). Curran Associates, Inc.
Åpne denne publikasjonen i ny fane eller vindu >>Convex formulations for training two-layer ReLU neural networks
2025 (engelsk)Inngår i: 13th International Conference on Learning Representations, ICLR 2025, Curran Associates, Inc., 2025, s. 30682-30704Konferansepaper, Publicerat paper (Fagfellevurdert)
Abstract [en]

Solving non-convex, NP-hard optimization problems is crucial for training machine learning models, including neural networks. However, non-convexity often leads to black-box machine learning models with unclear inner workings. While convex formulations have been used for verifying neural network robustness, their application to training neural networks remains less explored. In response to this challenge, we reformulate the problem of training infinite-width two-layer ReLU networks as a convex completely positive program in a finite-dimensional (lifted) space. Despite the convexity, solving this problem remains NP-hard due to the complete positivity constraint. To overcome this challenge, we introduce a semidefinite relaxation that can be solved in polynomial time. We then experimentally evaluate the tightness of this relaxation, demonstrating its competitive performance in test accuracy across a range of classification tasks.

sted, utgiver, år, opplag, sider
Curran Associates, Inc., 2025
Emneord
copositive programming, semidefinite programming, neural networks
HSV kategori
Identifikatorer
urn:nbn:se:umu:diva-236599 (URN)2-s2.0-105010230817 (Scopus ID)979-8-3313-2085-0 (ISBN)
Konferanse
International Conference on Learning Representations (ICLR), Singapore, April 24-28, 2025
Forskningsfinansiär
Wallenberg AI, Autonomous Systems and Software Program (WASP)Knut and Alice Wallenberg FoundationSwedish Research Council
Tilgjengelig fra: 2025-03-17 Laget: 2025-03-17 Sist oppdatert: 2025-07-18bibliografisk kontrollert
Dadras, A., Banerjee, S., Prakhya, K. & Yurtsever, A. (2024). Federated Frank-Wolfe algorithm. In: Albert Bifet; Jesse Davis; Tomas Krilavičius; Meelis Kull; Eirini Ntoutsi; Indrė Žliobaitė (Ed.), Machine learning and knowledge discovery in databases. Research track: European Conference, ECML PKDD 2024, Vilnius, Lithuania, September 9–13, 2024, proceedings, part III. Paper presented at European Conference on Machine Learning and Principles and Practice of Knowledge Discovery in Databases (ECML PKDD 2024), Vilnius, Lithuania, September 9-13, 2024 (pp. 58-75). Springer Nature
Åpne denne publikasjonen i ny fane eller vindu >>Federated Frank-Wolfe algorithm
2024 (engelsk)Inngår i: Machine learning and knowledge discovery in databases. Research track: European Conference, ECML PKDD 2024, Vilnius, Lithuania, September 9–13, 2024, proceedings, part III / [ed] Albert Bifet; Jesse Davis; Tomas Krilavičius; Meelis Kull; Eirini Ntoutsi; Indrė Žliobaitė, Springer Nature, 2024, s. 58-75Konferansepaper, Publicerat paper (Fagfellevurdert)
Abstract [en]

Federated learning (FL) has gained a lot of attention in recent years for building privacy-preserving collaborative learning systems. However, FL algorithms for constrained machine learning problems are still limited, particularly when the projection step is costly. To this end, we propose a Federated Frank-Wolfe Algorithm (FedFW). FedFW features data privacy, low per-iteration cost, and communication of sparse signals. In the deterministic setting, FedFW achieves an ε-suboptimal solution within O(ε-2) iterations for smooth and convex objectives, and O(ε-3) iterations for smooth but non-convex objectives. Furthermore, we present a stochastic variant of FedFW and show that it finds a solution within O(ε-3) iterations in the convex setting. We demonstrate the empirical performance of FedFW on several machine learning tasks.

sted, utgiver, år, opplag, sider
Springer Nature, 2024
Serie
Lecture Notes in Computer Science, ISSN 0302-9743, E-ISSN 1611-3349 ; 14943
Emneord
federated learning, frank wolfe, conditional gradient method, projection-free, distributed optimization
HSV kategori
Identifikatorer
urn:nbn:se:umu:diva-228614 (URN)10.1007/978-3-031-70352-2_4 (DOI)001308375900004 ()978-3-031-70351-5 (ISBN)978-3-031-70352-2 (ISBN)
Konferanse
European Conference on Machine Learning and Principles and Practice of Knowledge Discovery in Databases (ECML PKDD 2024), Vilnius, Lithuania, September 9-13, 2024
Forskningsfinansiär
Wallenberg AI, Autonomous Systems and Software Program (WASP)Swedish Research Council, 2023-05476
Merknad

Also part of the book sub series: Lecture Notes in Artificial Intelligence (LNAI). 

Tilgjengelig fra: 2024-08-19 Laget: 2024-08-19 Sist oppdatert: 2025-04-24bibliografisk kontrollert
Dadras, A., Prakhya, K. & Yurtsever, A. (2022). Federated Frank-Wolfe Algorithm. In: : . Paper presented at FL-NeurIPS'22, International Workshop on Federated Learning: Recent Advances and New Challenges in Conjunction with NeurIPS 2022, New Orleans, LA, USA, December 2, 2022.
Åpne denne publikasjonen i ny fane eller vindu >>Federated Frank-Wolfe Algorithm
2022 (engelsk)Konferansepaper, Poster (with or without abstract) (Fagfellevurdert)
Abstract [en]

Federated learning (FL) has gained much attention in recent years for building privacy-preserving collaborative learning systems. However, FL algorithms for constrained machine learning problems are still very limited, particularly when the projection step is costly. To this end, we propose a Federated Frank-Wolfe Algorithm (FedFW). FedFW provably finds an ε-suboptimal solution of the constrained empirical risk-minimization problem after O(ε−2) iterations if the objective function is convex. The rate becomes O(ε−3) if the objective is non-convex. The method enjoys data privacy, low per-iteration cost and communication of sparse signals. We demonstrate empirical performance of the FedFW algorithm on several machine learning tasks.

Emneord
federated learning, frank wolfe, conditional gradient method, projection-free, distributed optimization
HSV kategori
Identifikatorer
urn:nbn:se:umu:diva-205126 (URN)
Konferanse
FL-NeurIPS'22, International Workshop on Federated Learning: Recent Advances and New Challenges in Conjunction with NeurIPS 2022, New Orleans, LA, USA, December 2, 2022
Forskningsfinansiär
Wallenberg AI, Autonomous Systems and Software Program (WASP)
Tilgjengelig fra: 2023-02-23 Laget: 2023-02-23 Sist oppdatert: 2025-01-23bibliografisk kontrollert
Organisasjoner