Talaria: Interactively Optimizing Machine Learning Models for Efficient Inference

Hohman, Fred; Wang, Chaoqun; Lee, Jinmook; Görtler, Jochen; Moritz, Dominik; Bigham, Jeffrey P; Ren, Zhile; Foret, Cecile; Shan, Qi; Zhang, Xiaoyi

doi:10.1145/3613904.3642628

Computer Science > Human-Computer Interaction

arXiv:2404.03085 (cs)

[Submitted on 3 Apr 2024]

Title:Talaria: Interactively Optimizing Machine Learning Models for Efficient Inference

Authors:Fred Hohman, Chaoqun Wang, Jinmook Lee, Jochen Görtler, Dominik Moritz, Jeffrey P Bigham, Zhile Ren, Cecile Foret, Qi Shan, Xiaoyi Zhang

View PDF HTML (experimental)

Abstract:On-device machine learning (ML) moves computation from the cloud to personal devices, protecting user privacy and enabling intelligent user experiences. However, fitting models on devices with limited resources presents a major technical challenge: practitioners need to optimize models and balance hardware metrics such as model size, latency, and power. To help practitioners create efficient ML models, we designed and developed Talaria: a model visualization and optimization system. Talaria enables practitioners to compile models to hardware, interactively visualize model statistics, and simulate optimizations to test the impact on inference metrics. Since its internal deployment two years ago, we have evaluated Talaria using three methodologies: (1) a log analysis highlighting its growth of 800+ practitioners submitting 3,600+ models; (2) a usability survey with 26 users assessing the utility of 20 Talaria features; and (3) a qualitative interview with the 7 most active users about their experience using Talaria.

Comments:	Proceedings of the 2024 ACM CHI Conference on Human Factors in Computing Systems
Subjects:	Human-Computer Interaction (cs.HC); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2404.03085 [cs.HC]
	(or arXiv:2404.03085v1 [cs.HC] for this version)
	https://doi.org/10.48550/arXiv.2404.03085
Related DOI:	https://doi.org/10.1145/3613904.3642628

Submission history

From: Fred Hohman [view email]
[v1] Wed, 3 Apr 2024 21:55:44 UTC (19,385 KB)

Computer Science > Human-Computer Interaction

Title:Talaria: Interactively Optimizing Machine Learning Models for Efficient Inference

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Human-Computer Interaction

Title:Talaria: Interactively Optimizing Machine Learning Models for Efficient Inference

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators