Conceptio
›
parallelcomputing
Topic
parallelcomputing
Knowledge-graph topic
· documents ABOUT parallelcomputing across the archive
164
Documents about parallelcomputing
Documents about parallelcomputing
Verifying In-Network Computing Systems for Design Risks
#13074
arXiv CS
OffloadFS: Leveraging Disaggregated Storage for Computation Offloading
#13996
arXiv CS
Self-adaptive Multi-Access Edge Architectures: A Robotics Case
#13997
arXiv CS
PackSELL: A Sparse Matrix Format for Precision-Agnostic High-Performance SpMV
#13998
arXiv CS
Event Tensor: A Unified Abstraction for Compiling Dynamic Megakernel
#13999
arXiv CS
An Engineering Journey Training Large Language Models at Scale on Alps: The Apertus Experience
#14000
arXiv CS
Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems
#18979
arXiv CS
Scepsy: Serving Agentic Workflows Using Aggregate LLM Pipelines
#18980
arXiv CS
Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter
#18981
arXiv CS
Efficient calculation of available space for multi-NUMA virtual machines
#18982
arXiv CS
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
#18983
arXiv CS
Cooperate to Compete: Strategic Data Generation and Incentivization Framework for Coopetitive Cross-Silo Federated Learning
#18984
arXiv CS
Exploiting Correlations in Federated Learning: Opportunities and Practical Limitations
#18985
arXiv CS
ELMoE-3D: Leveraging Intrinsic Elasticity of MoE for Hybrid-Bonding-Enabled Self-Speculative Decoding in On-Premises Serving
#18986
arXiv CS
AgileLog: A Forkable Shared Log for Agents on Data Streams
#18987
arXiv CS
CoCoDiff: Optimizing Collective Communications for Distributed Diffusion Transformer Inference Under Ulysses Sequence Parallelism
#18988
arXiv CS
Fast Concurrent Primitives Despite Contention
#18989
arXiv CS
Parallel R-tree-based Spatial Query Processing on a Commercial Processing-in-Memory System
#18990
arXiv CS
Distributed Variational Quantum Linear Solver
#18991
arXiv CS
Incidence Constraints in Hypergraph Partitioning on GPU
#18992
arXiv CS
Training Time Prediction for Mixed Precision-based Distributed Training
#31237
arXiv CS
Logarithmic-Time Geodesically Convex Decomposition in Programmable Matter
#31238
arXiv CS
Compositional Design, Implementation, and Verification of Swarms (Technical Report)
#31239
arXiv CS
Robust Synchronisation for Federated Learning in The Face of Correlated Device Failure
#31240
arXiv CS
T-RBFT: A Scalable and Efficient Byzantine Consensus Based on Trusted Execution Environment for Consortium Blockchain
#31241
arXiv CS
Evaluating SYCL as a Unified Programming Model for Heterogeneous Systems
#31242
arXiv CS
Continuous benchmarking: Keeping pace with an evolving ecosystem of models and technologies
#31243
arXiv CS
New Kids: An Architecture and Performance Investigation of Second-Generation Serverless Platforms
#31244
arXiv CS
Breaking the Training Barrier of Billion-Parameter Universal Machine Learning Interatomic Potentials
#31245
arXiv CS
CroSatFL: Energy-Efficient Federated Learning with Cross-Aggregation for Satellite Edge Computing
#31246
arXiv CS
cuNNQS-SCI: A Fully GPU-Accelerated Framework for High-Performance Configuration Interaction Selection withNeural Network QQantum States
#31247
arXiv CS
Accuracy Is Speed: Towards Long-Context-Aware Routing for Distributed LLM Serving
#31248
arXiv CS
BlockRaFT: A Distributed Framework for Fault-Tolerant and Scalable Blockchain Nodes
#31249
arXiv CS
DataCenterGym: A Physics-Grounded Simulator for Multi-Objective Data Center Scheduling
#31250
arXiv CS
Optimizing Stochastic Gradient Push under Broadcast Communications
#31251
arXiv CS
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
#120464
arXiv CS
User Experiences with MPI RMA and ULFM in a Resilient Key-Value Store Implementation
#120465
arXiv CS
Trust, but Verify: ByzTwin-Range, a Digital Twin Cyber-Range for Byzantine Faults
#120466
arXiv CS
Optimizing Memory Allocation in Distributed Clusters with Predictive Modeling
#120467
arXiv CS
Toward Optimality: A Tighter Analysis of Message Complexity for Leader Election in Diameter-Two Networks
#120468
arXiv CS
Matrix-Free 3D SIMP Topology Optimization with Fused Gather-GEMM-Scatter Kernels
#120469
arXiv CS
GPUOS: A GPU Operating System Primitive for Transparent Operation Fusion
#120470
arXiv CS
AsyncSparse: Accelerating Sparse Matrix-Matrix Multiplication on Asynchronous GPU Architectures
#120471
arXiv CS
DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models
#120472
arXiv CS
Towards Energy Efficient Co-Scheduling in HPC
#120473
arXiv CS
EcoShift: Performance-Aware Power Management for Power-Constrained Heterogeneous Systems
#120474
arXiv CS
SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving
#120475
arXiv CS
Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML
#120476
arXiv CS
Active Inference-Based Adaptive Routing for Heterogeneous Edge AI Services
#120477
arXiv CS
Taming GPU Underutilization via Static Partitioning and Fine-grained CPU Offloading
#2534
arXiv CS
← Previous
Page 3 of 4
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.