SBAC-PAD 2026: 38TH IEEE/SBC INTERNATIONAL SYMPOSIUM ON COMPUTER ARCHITECTURE AND HIGH PERFORMANCE COMPUTING (SBAC-PAD)
TALK KEYWORD INDEX

This page contains an index consisting of author-provided keywords.

3
3d-stacked memory
A
AI accelerator
AI Inference
Allgather
AllReduce
Alltoall
Approximate Computing
architecture-aware optimization
Arm SPE
artificial intelligence
Automatic Data Distribution
B
bandwidth
Benchmarking
BFloat16
big data
bitstream cache
BLAS
Branch prediction
C
cache
cache coherence
cache multi-GPU inference MFU
Cache performance
Candidate Reduction
categorization
cloud economics
Cloud workload characterization
code generation
codesign
Collective Operations
Compiler infrastructures
compute marketplace
Computer Architecture
Convolution
CPU usage prediction
Cross-application interference
Cross-Layer Analysis
cross-vendor
CUDA
custom instruction
Cybersecurity
D
Data Augmentation
Data Pipelines
Data prefetching
Data synchronization
DDR5
decentralised infrastructure
Deep learning
deep learning parallelism SIMD AVX-512
Deep Learning Systems
Deep neural networks
Design Space Exploration
directory compression
Distributed Learning
DLFloat
DNN
DNNs
DRAM Security
Dynamic Frequency Scaling
dynamic resource management
E
eBPF
Edge AI
eFPGA
Emerging Memory Technologies
End-to-End Latency
Energy efficiency
epidemiology
Explainability
F
FARIMA
Fast Fourier Transform
Feature Pyramid Networks
Field-Programmable Gate Array (FPGA)
FMA data parallelism tensor parallelism NVLink all-reduce
FPGA
FPGAs
Fractional Brownian Motion
G
GEMM
Geophysical Inverse Problems
GFLOPS
Ghost Cell Halo Exchange
GPU
GPU acceleration
GPU Architecture
GPU computing
GPU Partitioning
GPU Sharing
GPU Utilization
Green HPC
H
Hardware-Software Co-Design
Heterogeneous Memory Systems
high performance
High performance computing
High-Performance Computing
High-performance computing (HPC)
HIP
HPC
HPC job
hyperspectral imaging
I
I/O
Im2col
indoor localization
industrial inspection
Instruction-Level Parallelism
interference
Interference profiling
Interference-aware scheduling
Intrusion detection
Irregular and graph workloads
K
kernel profiling
L
Large scale simulations
latency-constrained systems
Leaky Integrate and Fire (LIF)
linear algebra
Long-range dependence
M
malleability
Matrix Multiplication
Medical Image Analysis
memory bandwidth roofline Amdahl’s law out-of-memory KV-
Memory Controller
micro-kernels
Microarchitecture simulation
mixed-criticality
MLIR
MLP accelerator
MNIST
Model Compression
monitoring
MPI
MPI Domain Decomposition
multi-GPU
multicore
N
near-data processing
Neighborhood Collective
NSL-KDD dataset
Numerical Stability
O
Object Detection
Out-of-Order Execution
P
parallel algorithms
Parallelization
pattern
Performance
performance and scaling analysis
Performance Evaluation
PRAC
Proactive VM migration
Programmable Logic
PyTorch
Q
Q-function
Quantum Machine Learning
R
real-time
real-time processing
reconfigurable instructions
Reduced Precision
Reinforcement Learning
Resource management
RISC-V
RISC-V Vector extension
ROCm
RowHammer
Runtime prediction
S
SBCs
Scalable Computing
Seismic RTM (Reverse Time Migration)
Seismic wave propagation
Shared Memory
short-time Fourier transform
SIMD vectorization
SLA compliance
Slurm
soft instruction
Solid-State Drives(SSD)
sparse directory
Spectral-finite-element method
Spiking Neural Network (SNN)
Spot instances
stablecoin
Storage Systems
Sustainable Software
Synthetic Data Generation
System-level instrumentation
SystemTap
T
taxonomy
token escrow
U
Uncertainty Estimation
Upper-bound analysis
V
Value-at-Risk
Virtual Topology
W
walltime
Warp Scheduling
WiFi finger-printing
Workload characterization
X
xDSL