Conceptio
›
parallel-computing
Topic
parallel-computing
Knowledge-graph topic
· documents ABOUT parallel-computing across the archive
1,734
Documents about parallel-computing
Documents about parallel-computing
Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents
#329051
arXiv CS
ROSA: A Robotics Foundation Model Serving System for Robot Factories
#329052
arXiv CS
Taming GPU Underutilization via Static Partitioning and Fine-grained CPU Offloading
#2534
arXiv CS
ALTO: Adaptive LoRA Tuning and Orchestration for Heterogeneous LoRA Training Workloads
#2569
arXiv CS
Edge-Oriented Orchestration of Energy Services Using Graph-Driven Swarm Intelligence
#2577
arXiv CS
NBI-Slurm: Simplified submission of Slurm jobs with energy saving mode
#2580
arXiv CS
Sustaining Exascale Performance: Lessons from HPL and HPL-MxP on Aurora
#5933
arXiv CS
ALTO: Adaptive LoRA Tuning and Orchestration for Heterogeneous LoRA Training Workloads
#5942
arXiv CS
FEDBUD: Joint Incentive and Privacy Optimization for Resource-Constrained Federated Learning
#10317
arXiv CS
Evaluating Cross-Architecture Performance Modeling of Distributed ML Workloads Using StableHLO
#13066
arXiv CS
Understanding Large-Scale HPC System Behavior Through Cluster-Based Visual Analytics
#13070
arXiv CS
PackSELL: A Sparse Matrix Format for Precision-Agnostic High-Performance SpMV
#13998
arXiv CS
Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems
#18979
arXiv CS
BlockRaFT: A Distributed Framework for Fault-Tolerant and Scalable Blockchain Nodes
#31249
arXiv CS
DataCenterGym: A Physics-Grounded Simulator for Multi-Objective Data Center Scheduling
#31250
arXiv CS
Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML
#120476
arXiv CS
Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling
#120478
arXiv CS
From Swap Axioms to Weighted Geometric Means: A Characterization of AMMs
#120485
arXiv CS
LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing
#124015
arXiv CS
Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
#124029
arXiv CS
Ocean: Fast Estimation-Based Sparse General Matrix-Matrix Multiplication on GPU
#124030
arXiv CS
Quantum-HPC Software Stacks and the openQSE Reference Architecture: A Survey
#126493
arXiv CS
FreeScale: Distributed Training for Sequence Recommendation Models with Minimal Scaling Cost
#138896
arXiv CS
FlashOverlap: Minimizing Tail Latency in Communication Overlap for Distributed LLM Training
#138898
arXiv CS
Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities
#138905
arXiv CS
SpecFed: Accelerating Federated LLM Inference with Speculative Decoding and Compressed Transmission
#141454
arXiv CS
Enhancing Performance Insight at Scale: A Heterogeneous Framework for Exascale Diagnostics
#155228
arXiv CS
Distributed Observer-based Fault Detection over Intelligent Networked Multi-Vehicle Systems
#155241
arXiv CS
Cross-Layer Energy Analysis of Multimodal Training on Grace Hopper Superchips
#155249
arXiv CS
CvxCluster: Solving Large, Complex, Granular Resource Allocation Problems 100-1000x Faster
#155255
arXiv CS
A Domain-Driven Design Simulator for Business Logic-Rich Microservice Systems
#155259
arXiv CS
SURGE: SuperBatch Unified Resource-efficient GPU Encoding for Heterogeneous Partitioned Data
#155262
arXiv CS
ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL
#168262
arXiv CS
A Scalable Digital Twin Framework for Energy Optimization in Data Centers
#168279
arXiv CS
Privacy-preserving Chunk Scheduling in a BitTorrent Implementation of Federated Learning
#175197
arXiv CS
Cloud Performance Decomposition for Long-Term Performance Engineering: A Case Study
#175208
arXiv CS
KV-RM: Regularizing KV-Cache Movement for Static-Graph LLM Serving
#175210
arXiv CS
TS-Verkle: A TypeScript Native Verkle Library With On-chain Verifier
#175223
arXiv CS
Unleashing Scalable Context Parallelism for Foundation Models Pre-Training via FCP
#175229
arXiv CS
NCCLZ: Compression-Enabled GPU Collectives with Decoupled Quantization and Entropy Coding
#178841
arXiv CS
Trade-offs in Decentralized Agentic AI Discovery Across the Compute Continuum
#178844
arXiv CS
Position: LLM Inference Should Be Evaluated as Energy-to-Token Production
#178845
arXiv CS
DisagMoE: Computation-Communication overlapped MoE Training via Disaggregated AF-Pipe Parallelism
#178857
arXiv CS
Rescaled Asynchronous SGD: Optimal Distributed Optimization under Data and System Heterogeneity
#180625
arXiv CS
Efficient and Portable Support for Overdecomposition on Distributed Memory GPGPU Platforms
#180634
arXiv CS
PCDM: A Diffusion-Based Data Poisoning Attack Against Federated Learning Systems
#194283
arXiv CS
Mosaic: Towards Efficient Training of Multimodal Models with Spatial Resource Multiplexing
#200441
arXiv CS
The Task Completion Problem and its Application to Crash-Resilient Computation
#200457
arXiv CS
TierCheck: Tiered Checkpointing for Fault Tolerance in Large Language Model Training
#200461
arXiv CS
HexAGenT: Efficient Agentic LLM Serving via Workflow- and Heterogeneity-Aware Scheduling
#200468
arXiv CS
← Previous
Page 10 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.