Conceptio
›
and-cluster-computing
Topic
and-cluster-computing
Knowledge-graph topic
· documents ABOUT and-cluster-computing across the archive
656
Documents about and-cluster-computing
Documents about and-cluster-computing
Optimizing Allreduce Operations for Modern Heterogeneous Architectures with Multiple Processes per GPU
#670369
arXiv (All)
Semantic Lock: Synchronization Based on the Analysis of the Operation Conflict Graph
#670649
arXiv (All)
PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand Response
#789869
arXiv (OAI Expanded)
PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand Response
#793888
arXiv (All)
Projection-Free Bandit Online Optimization for Multi-Agent Systems with Dynamic Regret
#797959
arXiv (OAI Expanded)
Projection-Free Bandit Online Optimization for Multi-Agent Systems with Dynamic Regret
#799382
arXiv (All)
mzCache: On-Device LLM Memory Management under Multitasking
#821283
arXiv (All)
RT-HiSS: Ray Tracing Accelerated High Dimensional Vector Similarity Searches
#822386
arXiv (All)
AceSpec: An Asymmetric Edge-Cloud Collaborative Framework for Communication-Efficient LLM Inference
#920443
arXiv (OAI Expanded)
AceSpec: An Asymmetric Edge-Cloud Collaborative Framework for Communication-Efficient LLM Inference
#924677
arXiv (All)
BASP: Communication-Efficient Batch-Aware Sequence Parallelism for LLM Training
#925514
arXiv (All)
JuPyLive: Seamless Migration of Jupyter Notebook Resources from Laptop to HPC
#926471
arXiv (All)
Atlas: Optimizing Deployment of Compound AI Workflows on Heterogeneous Clusters
#927797
arXiv (All)
GreenPipe: Power Modeling for Containerized DNN Inference on Kubernetes Edge Nodes
#971917
arXiv (All)
Speculation at a Distance: Where Edge-Cloud Speculative Decoding Actually Pays Off
#972453
arXiv (All)
Poseidon: DAG-Guided Parallelism Search for LLM Pre-Training on Heterogeneous Clusters
#973693
arXiv (All)
Kino-PAX$^+$: Near-Optimal Massively Parallel Kinodynamic Sampling-based Motion Planner
#987581
arXiv (All)
Breaking Fault Lines: Unifying TEE-Assisted BFT Consensus in Partially Trusted Worlds
#996567
arXiv (All)
Towards Reproducible Evaluation of Distributed Quantum Circuit Partitioning Algorithms
#996939
arXiv (All)
CEDD-optimizer: Enabling Cost-Efficient Dataset Distillation on Geographically Distributed Edge Systems
#997062
arXiv (All)
Gutenberg: Taming Latency-Critical Cloud Services with Near-Data-Processing
#999889
arXiv (All)
SpecReuse: Spectral Graph Reuse for Efficient Vision GNN Inference on FPGAs
#1011443
arXiv (All)
Token Latency Fairness: Performance Isolation for Multi-Tenant LLM Serving
#1011811
arXiv (All)
PipeSwift: Revisiting Pipeline Parallelism for Large-Scale Completion-Oriented Agentic LLM Serving
#1012381
arXiv (All)
From Pixels to Semantics: Edge AI for UAV-Based Critical Infrastructure Inspection
#1012633
arXiv (All)
VERA: Reinforcement Learning for Dynamic Memory Scaling of HPC Workloads in Kubernetes
#1014981
arXiv (All)
Accelerating Sharded Data Parallelism at Scale with Federated Learning
#1034123
arXiv (All)
Profiling Concurrent Vision Inference Workloads on NVIDIA Jetson -- Extended
#1036564
arXiv (All)
A Heavy-Load-Enhanced and Changeable-Periodicity-Perceived Workload Prediction Network
#139055
arXiv (OAI)
SpecBranch: Speculative Decoding via Hybrid Drafting and Rollback-Aware Branch Parallelism
#141752
arXiv (OAI)
FeLoG: Scalable and Efficient Distributed Graph Embedding with Feedback Loop Mechanism
#650908
arXiv (OAI Expanded)
FeLoG: Scalable and Efficient Distributed Graph Embedding with Feedback Loop Mechanism
#652361
arXiv (All)
Characterizing the I/O Behavior of HPC Applications through Modeling and Simulation
#779567
arXiv (OAI Expanded)
Characterizing the I/O Behavior of HPC Applications through Modeling and Simulation
#780942
arXiv (All)
Failure-Resilient and Carbon-Efficient Deployment of Microservices over the Cloud-Edge Continuum
#797302
arXiv (OAI Expanded)
Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption
#797486
arXiv (OAI Expanded)
Failure-Resilient and Carbon-Efficient Deployment of Microservices over the Cloud-Edge Continuum
#798716
arXiv (All)
Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption
#798900
arXiv (All)
Characterizing the Scalability and Performance of Large-Scale AI Training Under Multi-Tenancy
#926227
arXiv (All)
Unified AI Gateway: A Framework for Joint Model Routing and KV Cache Management
#975016
arXiv (All)
The Ghost in the Datacenter: Link Flapping, Topology Knowledge Failures, and the FITO Category Mistake
#987603
arXiv (All)
Serving Agentic Workflows with a Physical-Plan Compiler and Adaptive Runtime
#1011154
arXiv (All)
Replication-Aware Placement of Functions and Data in the Edge-Cloud Continuum
#1013089
arXiv (All)
Hopper: Bounded-Memory Collaborative Debiasing for Byzantine-Tolerant Peer Sampling
#1035649
arXiv (All)
SkelOT: Reusing AOT Compilation Across EVM Contract Families
#1050583
arXiv (All)
PRISM: Probabilistic Runtime Insights and Scalable Performance Modeling for Large-Scale Distributed Training
#1051431
arXiv (All)
Issues in petabyte data indexing, retrieval and analysis
#401908
DataCite
Decentralized Nonconvex Optimization under Heavy-Tailed Noise: Normalization and Optimal Convergence
#147084
arXiv (OAI)
Characterising Global Platforms: Centralised, Decentralised, Federated, and Grassroots
#789502
arXiv (OAI Expanded)
Characterising Global Platforms: Centralised, Decentralised, Federated, and Grassroots
#793506
arXiv (All)
← Previous
Page 6 of 16
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.