Conceptio
›
parallel-computing
Topic
parallel-computing
Knowledge-graph topic
· documents ABOUT parallel-computing across the archive
1,734
Documents about parallel-computing
Documents about parallel-computing
Accuracy Is Speed: Towards Long-Context-Aware Routing for Distributed LLM Serving
#31248
arXiv CS
Matrix-Free 3D SIMP Topology Optimization with Fused Gather-GEMM-Scatter Kernels
#120469
arXiv CS
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
#120475
arXiv CS
GreenPeas: Unlocking Adaptive Quantum Error Correction with Just-in-Time Decoding Hypergraphs
#120490
arXiv CS
FedSIR: Spectral Client Identification and Relabeling for Federated Learning with Noisy Labels
#124006
arXiv CS
Distributed Quantum-Enhanced Optimization: A Topographical Preconditioning Approach for High-Dimensional Search
#124008
arXiv CS
Distributed Quantum Optimization for Large-Scale Higher-Order Problems with Dense Interactions
#124009
arXiv CS
FASER: Fine-Grained Phase Management for Speculative Decoding in Dynamic LLM Serving
#124010
arXiv CS
Predictive Autoscaling for Node.js on Kubernetes: Lower Latency, Right-Sized Capacity
#124017
arXiv CS
Optimal Routing for Federated Learning over Dynamic Satellite Networks: Tractable or Not?
#124022
arXiv CS
Mass Matrix Assembly on Tensor Cores for Implicit Particle-In-Cell Methods
#124025
arXiv CS
GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGA
#126485
arXiv CS
Optimizing High-Throughput Distributed Data Pipelines for Reproducible Deep Learning at Scale
#126486
arXiv CS
Data-Free Contribution Estimation in Federated Learning using Gradient von Neumann Entropy
#134520
arXiv CS
$O(K)$-Approximation Coflow Scheduling in $K$-Core Optical Circuit Switching Networks
#134525
arXiv CS
Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning
#134528
arXiv CS
KubePACS: Kubernetes Cluster Using Performant, Highly Available, and Cost Efficient Spot Instances
#138897
arXiv CS
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
#138901
arXiv CS
Cloud to Edge: Benchmarking LLM Inference On Hardware-Accelerated Single-Board Computers
#141464
arXiv CS
Caliper-in-the-Loop: Black-Box Optimization for Hyperledger Fabric Performance Tuning
#155238
arXiv CS
Joint Temporal-Structural Representation Learning for Distributed Fault Discrimination in Microservice Architectures
#155253
arXiv CS
Position: LLM Serving Needs Mathematical Optimization and Algorithmic Foundations, Not Just Heuristics
#155258
arXiv CS
LLM-Emu: Native Runtime Emulation of LLM Inference via Profile-Driven Sampling
#155264
arXiv CS
AGoQ: Activation and Gradient Quantization for Memory-Efficient Distributed Training of LLMs
#155265
arXiv CS
Token Arena: A Continuous Benchmark Unifying Energy and Cognition in AI Inference
#155267
arXiv CS
A Privacy-Preserving Machine Learning Framework for Edge Intelligence: An Empirical Analysis
#168274
arXiv CS
One Pool, Two Caches: Adaptive HBM Partitioning for Accelerating Generative Recommender Serving
#168287
arXiv CS
Revocation-Ready CP-ABE Key Management for Blockchain-Based IoT Data Sharing
#168292
arXiv CS
Closer in the Gap: Towards Portable Performance on RISC-V Vector Processors
#175193
arXiv CS
Accelerating Locality-Driven Integration in Quantum Chemistry with Block-Structured Matrix Multiplication
#175202
arXiv CS
Lakestream: A Consistent and Brokerless Data Plane for Large Foundation Model Training
#175205
arXiv CS
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
#175215
arXiv CS
Accelerating Precise End-to-End Simulation: Latency-Sensitive Many-core System Modeling
#175236
arXiv CS
MLCommons Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces
#178853
arXiv CS
PipeSD: An Efficient Cloud-Edge Collaborative Pipeline Inference Framework with Speculative Decoding
#180630
arXiv CS
An efficient multi-GPU implementation for the Discontinuous Galerkin ocean model SLIM
#194284
arXiv CS
ParamSpMM: Adaptive and Efficient Sparse Matrix-Matrix Multiplication on GPUs for GNNs
#194289
arXiv CS
Duet instrumentation: An Agentic Approach to Improving Sensitivity in Cloud Service Benchmarking
#200450
arXiv CS
Guard: Scalable Straggler Detection and Node Health Management for Large-Scale Training
#200460
arXiv CS
OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization
#200462
arXiv CS
S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination
#200465
arXiv CS
GoodServe: Towards High-Goodput Serving of Agentic LLM Inferences over Heterogeneous Resources
#200466
arXiv CS
Nf-PEAK: Process-Based Energy Attribution for Nextflow Workflows on Kubernetes Clusters
#216783
arXiv CS
LatentBox: Storing AI-Generated Images at Scale via a Latent-First Design
#216805
arXiv CS
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
#224469
arXiv CS
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
#224479
arXiv CS
Multi-Factor Trust-Driven Secure Communication Model for Cloud-Based Digital Twins
#224486
arXiv CS
Multi-Round Visibility: A Post-Consensus Ordering Layer for DAG-Based BFT
#224487
arXiv CS
Design and Implementation of a Serverless MapReduce Framework for Scalable Data Pipelines
#238567
arXiv CS
Understanding and Reducing Metadata-Driven Host Overheads in Sampling-Based GNN Training
#238569
arXiv CS
← Previous
Page 13 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.