Latest Research in Parallel Computing and Optimization Techniques

22 research papers · 2026 median publication year

Top Research Topics in Parallel Computing and Optimization Techniques

Highest-Cited Papers

  1. Evaluating OpenMP Offloading for Intra-node Multi-GPU Programming across NVIDIA, AMD, and Intel Architectures: A 3D Heat Transfer Case Study
  2. Cnuas: A Software-Defined AI/HPC Rack-scale Emulation Platform and Hyperscale Data Center Facility Twin
  3. Tools-CC-Bench: a Benchmark Suite for Collective Communication with Compression in HPC and AI Workloads
  4. Research on MuE_shui Effective Computing Efficiency System Construction, Hardware-Software Decoupling and Global Intelligent Evaluation
  5. HPCRSE@RSECon26: 5th annual meeting of the HPC RSE community
  6. Making MPI Collective Operations Visible: Understanding Internal Algorithms and Performance Implications
  7. ORCHA: A Performance Portability System for Extreme Heterogeneity
  8. PASCAL: A Phase-Aware Shared-Cache Model for Parallel Scans
  9. Vectorizer: Vectorizing NumPy Programs with Shape-Guided Rewrite
  10. What Is Windows' Hardware-Accelerated GPU Scheduling? — Does Turning It On Make Your PC Faster? (archived 2026-09-07)
  11. What Is Windows' Hardware-Accelerated GPU Scheduling? — Does Turning It On Make Your PC Faster? (archived 2026-09-07)
  12. JLIR: A Julia-Native MLIR-Inspired Intermediate Representation with Automatic JACC Kernel Extraction
  13. Trust, but Verify: Rigorously Profiling Best-Effort High-Performance Computing for Digital Evolution
  14. ORCHA: A performance portability system for extreme heterogeneity
  15. HPC Carpentry
  16. Performance Characterization of SPEC CPU 2026 on AMD EPYC 9755 Processor
  17. LLM Inference on IMC-NoC Architecture with Balanced Dataflow and Fine-Grained Parallelism
  18. Native FreeToken Serving on AMD Strix Halo: A ROCm/HIP Port and Controlled Unified-Memory Evaluation
  19. When Schedulers Meet Converters: A Case for Energy-Readiness-Aware AI Workload Placement
  20. DGNA: Dissecting GPU NUMA Architecture through Microbenchmarking and Data Analysis
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
L3 Region - - 2026 Sep Q3

Parallel Computing and Optimization Techniques

22 papers

Top Topics (10)

Parallel Computing and Optimization Techniques7
Distributed, Parallel, and Cluster Computing4
Hardware Architecture2
Advanced Computing and Algorithms1
Scientific Computing and Data Management1
Numerical Analysis1
Performance1
Computation and Language1
Programming Languages1
Neural and Evolutionary Computing1

Top Publications (20)

1.Evaluating OpenMP Offloading for Intra-node Multi-GPU Programming across NVIDIA, AMD, and Intel Architectures: A 3D Heat Transfer Case Study2.Cnuas: A Software-Defined AI/HPC Rack-scale Emulation Platform and Hyperscale Data Center Facility Twin3.Tools-CC-Bench: a Benchmark Suite for Collective Communication with Compression in HPC and AI Workloads4.Research on MuE_shui Effective Computing Efficiency System Construction, Hardware-Software Decoupling and Global Intelligent Evaluation5.HPCRSE@RSECon26: 5th annual meeting of the HPC RSE community6.Making MPI Collective Operations Visible: Understanding Internal Algorithms and Performance Implications7.ORCHA: A Performance Portability System for Extreme Heterogeneity8.PASCAL: A Phase-Aware Shared-Cache Model for Parallel Scans9.Vectorizer: Vectorizing NumPy Programs with Shape-Guided Rewrite10.What Is Windows' Hardware-Accelerated GPU Scheduling? — Does Turning It On Make Your PC Faster? (archived 2026-09-07)11.What Is Windows' Hardware-Accelerated GPU Scheduling? — Does Turning It On Make Your PC Faster? (archived 2026-09-07)12.JLIR: A Julia-Native MLIR-Inspired Intermediate Representation with Automatic JACC Kernel Extraction13.Trust, but Verify: Rigorously Profiling Best-Effort High-Performance Computing for Digital Evolution14.ORCHA: A performance portability system for extreme heterogeneity15.HPC Carpentry16.Performance Characterization of SPEC CPU 2026 on AMD EPYC 9755 Processor17.LLM Inference on IMC-NoC Architecture with Balanced Dataflow and Fine-Grained Parallelism18.Native FreeToken Serving on AMD Strix Halo: A ROCm/HIP Port and Controlled Unified-Memory Evaluation19.When Schedulers Meet Converters: A Case for Energy-Readiness-Aware AI Workload Placement20.DGNA: Dissecting GPU NUMA Architecture through Microbenchmarking and Data Analysis
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.