FCoral: A Framework for Fine-grained Concurrency Evaluation Across Programming Language Runtimes

Modern programming languages increasingly support lightweight concurrency abstractions (e.g., coroutines), which are managed by language-level runtimes to improve scalability and performance in highly concurrent applications. Compared with coarse-grained thread-level concurrency, the performance of fine-grained coroutine-level concurrency depends not only on system resources but is even more sensitive to the efficiency of the concurrent runtimes. Performance evaluation across concurrent runtimes is important for understanding performance bottlenecks and improving the efficiency of both concurrent applications and runtime systems. However, existing approaches mainly rely on manually crafted, language-specific benchmarks, making it difficult to ensure semantic consistency across languages and perform fair cross-runtime comparisons. Moreover, they often treat concurrent runtimes as black boxes, overlooking the impact of internal runtime mechanisms on concurrency performance. In this paper, we propose FCoral, a cross-language framework for fine-grained concurrency evaluation. By constructing formally verified and language-agnostic Concurrent Workload Programs (CWPs), FCoral automatically generates semantically consistent benchmarks across different runtimes. It further employs a parameterized methodology to evaluate runtime performance across different concurrency scales, task granularities, and workload types. Using FCoral, we conduct a comprehensive study of five representative runtimes–—FFRT (C++), Go, JVM, Tokio (Rust), and Cangjie–—across CPU-intensive, I/O-intensive, and mixed workloads. We further evaluate FCoral on two real-world high-concurrency applications. Results show that FCoral successfully constructs cross-language benchmarks, facilitating systematic performance comparisons across concurrent runtimes. Based on the evaluation results, we propose two application-level optimization strategies. Experimental results demonstrate that these strategies not only achieve significant performance improvements over the unoptimized versions but also further validate the effectiveness of FCoral.

Authors

Institutions

Publication Details

Journal
ACM Transactions on Architecture and Code Optimization
Published
2026-10-03
DOI
https://doi.org/10.1145/3847668
Primary Topic
Parallel Computing and Optimization Techniques
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
OCT
article

FCoral: A Framework for Fine-grained Concurrency Evaluation Across Programming Language Runtimes

Weixing Ji, Jianhua Gao, Yuxiang Zhang, Danying Ge et al.
ACM Transactions on Architecture and Code Optimization
Parallel Computing and Optimization Techniques
article

FCoral: A Framework for Fine-grained Concurrency Evaluation Across Programming Language Runtimes

Weixing Ji, Jianhua Gao, Yuxiang Zhang, Danying Ge, Jianjun Shi, Bingxin Liu, Yinghui Huang
article en

Abstract

Modern programming languages increasingly support lightweight concurrency abstractions (e.g., coroutines), which are managed by language-level runtimes to improve scalability and performance in highly concurrent applications. Compared with coarse-grained thread-level concurrency, the performance of fine-grained coroutine-level concurrency depends not only on system resources but is even more sensitive to the efficiency of the concurrent runtimes. Performance evaluation across concurrent runtimes is important for understanding performance bottlenecks and improving the efficiency of both concurrent applications and runtime systems. However, existing approaches mainly rely on manually crafted, language-specific benchmarks, making it difficult to ensure semantic consistency across languages and perform fair cross-runtime comparisons. Moreover, they often treat concurrent runtimes as black boxes, overlooking the impact of internal runtime mechanisms on concurrency performance. In this paper, we propose FCoral, a cross-language framework for fine-grained concurrency evaluation. By constructing formally verified and language-agnostic Concurrent Workload Programs (CWPs), FCoral automatically generates semantically consistent benchmarks across different runtimes. It further employs a parameterized methodology to evaluate runtime performance across different concurrency scales, task granularities, and workload types. Using FCoral, we conduct a comprehensive study of five representative runtimes–—FFRT (C++), Go, JVM, Tokio (Rust), and Cangjie–—across CPU-intensive, I/O-intensive, and mixed workloads. We further evaluate FCoral on two real-world high-concurrency applications. Results show that FCoral successfully constructs cross-language benchmarks, facilitating systematic performance comparisons across concurrent runtimes. Based on the evaluation results, we propose two application-level optimization strategies. Experimental results demonstrate that these strategies not only achieve significant performance improvements over the unoptimized versions but also further validate the effectiveness of FCoral.

ACM Transactions on Architecture and Code Optimization
Beijing Normal University (CN)
Openalex Percentile: Top 6%
Parallel Computing and Optimization Techniques
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.