2023 IEEE 16th International Conference on Cloud Computing (CLOUD) 2023
DOI: 10.1109/cloud60044.2023.00021
|View full text |Cite
|
Sign up to set email alerts
|

IRIS: Interference and Resource Aware Predictive Orchestration for ML Inference Serving

Aggelos Ferikoglou,
Panos Chrysomeris,
Achilleas Tzenetopoulos
et al.

Abstract: Over the last years, the ever-growing number of Machine Learning(ML) and Artificial Intelligence(AI) applications deployed in the Cloud has led to high demands on the computing resources required for efficient processing. Multiple users deploy multiple applications on the same server node to maximize Quality of Service(QoS); however, this leads to increased interference. In addition, Cloud providers aim to minimize their operating costs by efficiently utilizing the available resources. These conflicting optimi… Show more

Help me understand this report

Search citation statements

Order By: Relevance

Paper Sections

Select...

Citation Types

0
0
0

Year Published

2024
2024
2024
2024

Publication Types

Select...
2
1

Relationship

0
3

Authors

Journals

citations
Cited by 3 publications
references
References 47 publications
0
0
0
Order By: Relevance