[论文解读] A Review of Serverless Use Cases and their Characteristics
本技术报告分析了来自四个数据源的89个无服务器用例,旨在描述一般特征、工作负载、应用和工作流特征,突出 AWS 的主导地位、生产部署情况以及典型用例。
The serverless computing paradigm promises many desirable properties for cloud applications - low-cost, fine-grained deployment, and management-free operation. Consequently, the paradigm has underwent rapid growth: there currently exist tens of serverless platforms and all global cloud providers host serverless operations. To help tune existing platforms, guide the design of new serverless approaches, and overall contribute to understanding this paradigm, in this work we present a long-term, comprehensive effort to identify, collect, and characterize 89 serverless use cases. We survey use cases, sourced from white and grey literature, and from consultations with experts in areas such as scientific computing. We study each use case using 24 characteristics, including general aspects, but also workload, application, and requirements. When the use cases employ workflows, we further analyze their characteristics. Overall, we hope our study will be useful for both academia and industry, and encourage the community to further share and communicate their use cases. This article appears also as a SPEC Technical Report: https://research.spec.org/fileadmin/user_upload/documents/rg_cloud/endorsed_publications/SPEC_RG_2020_Serverless_Usecases.pdf The article may be submitted for peer-reviewed publication.
研究动机与目标
- 从多源数据中识别并分类现实世界的无服务器用例。
- 表征无服务器工作负载的一般属性、工作负载、应用、需求与工作流特征。
- 评估无服务器用例的生产部署、开源可用性及领域分布。
- 提供洞见以指导平台设计、采用决策及未来研究。
提出的方法
- 系统性地收集来自开源项目、白皮书、灰色文献及科学计算来源的89个无服务器用例。
- 双评审对每个用例在24个预定义特征上进行特征描述。
- 计算评审者之间的一致性(Fleiss kappa)并整合有冲突的注释。
- 在公开的 Zenodo 仓库中托管完整数据集。
- 分类为一般、工作负载、应用、需求和工作流等类别。
实验结果
研究问题
- RQ1现实世界无服务器用例中常见的平台、应用领域和生产状态是什么?
- RQ2在实际应用中,哪些工作负载、应用和工作流特征可概括无服务器部署?
- RQ3开源可用性和跨领域适用性在无服务器用例中如何分布?
- RQ4组织采用无服务器的动机(成本、可扩展性、可维护性、性能)是什么,以及延迟/本地性要求出现的频率如何?
主要发现
- AWS 是主导的部署平台(80% 的用例)。
- Web 服务是最常见的应用领域(33%),其中55%的用例处于生产状态,40%为业务关键。
- API、流/异步处理、批处理任务,以及运维/监控是领先的应用类型(分别占28%、27%、23%和20%)。
- 55% 的分析用例处于生产环境;约36% 没有明确的生产证据,未知项在生产中被视为非生产。
- 53% 的用例是开源;47% 是闭源。
- 62% 的用例依赖存储或数据库;其他常见服务包括 API 网关和发布-订阅(pub-sub)。
- 81% 的工作负载属于突发型;大多数执行是按需而非定时;许多是轻量级的 HTTP 触发。
更好的研究,从现在开始
从阅读论文到最终审阅,大幅缩短您的研究时间。
无需绑定信用卡
本解读由 AI 生成,并经人工编辑审核。