Conceptio
›
cloud-computing
Topic
cloud-computing
Knowledge-graph topic
· documents ABOUT cloud-computing across the archive
170
Documents about cloud-computing
Documents about cloud-computing
Software-defined Cloud Manufacturing for Industry 4.0
#199546
OpenAlex
Wattlytics: A Web Platform for Co-Optimizing Performance, Energy, and TCO in HPC Clusters
#2539
arXiv CS
Blink: CPU-Free LLM Inference by Delegating the Serving Stack to GPU and SmartNIC
#2544
arXiv CS
Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning
#2546
arXiv CS
Beyond End-to-End: Dynamic Chain Optimization for Private LLM Adaptation on the Edge
#2556
arXiv CS
SwarmIO: Towards 100 Million IOPS SSD Emulation for Next-generation GPU-centric Storage Systems
#2558
arXiv CS
GTaP: A GPU-Resident Fork-Join Task-Parallel Runtime with a Pragma-Based Interface
#2565
arXiv CS
MegaTrain: Full Precision Training of 100B+ Parameter Large Language Models on a Single GPU
#2571
arXiv CS
CUTEv2: Unified and Configurable Matrix Extension for Diverse CPU Architectures with Minimal Design Overhead
#10308
arXiv CS
FlexVector: A SpMM Vector Processor with Flexible VRF for GCNs on Varying-Sparsity Graphs
#10324
arXiv CS
LoDAdaC: a unified local training-based decentralized framework with adaptive gradients and compressed communication
#10325
arXiv CS
An Engineering Journey Training Large Language Models at Scale on Alps: The Apertus Experience
#13054
arXiv CS
Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents
#13065
arXiv CS
An Engineering Journey Training Large Language Models at Scale on Alps: The Apertus Experience
#14000
arXiv CS
Parallel R-tree-based Spatial Query Processing on a Commercial Processing-in-Memory System
#18990
arXiv CS
User Experiences with MPI RMA and ULFM in a Resilient Key-Value Store Implementation
#120465
arXiv CS
A comprehensive evaluation of spatial co-execution on GPUs using MPS and MIG technologies
#134522
arXiv CS
Accelerating Intra-Node GPU-to-GPU Communication Through Multi-Path Transfers with CUDA Graphs
#134524
arXiv CS
SpotVista: Availability-Aware Recommendation System for Reliable and Cost-Efficient Multi-Node Spot Instances
#138890
arXiv CS
From Barrier to Bridge: The Case for AI Data Center/Power Grid Co-Design
#155232
arXiv CS
From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications
#155234
arXiv CS
Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon
#175211
arXiv CS
MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production
#175220
arXiv CS
Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction
#175221
arXiv CS
GriNNder: Breaking the Memory Capacity Wall in Full-Graph GNN Training with Storage Offloading
#178849
arXiv CS
The Distributed Complexity Landscape on Trees Depends on the Knowledge About the Network Size
#180633
arXiv CS
Charon: A Unified and Fine-Grained Simulator for Large-Scale LLM Training and Inference
#200464
arXiv CS
FedADAS: Communication-Efficient Federated Distillation for On-Device Driver Yawn Recognition in Vehicular Networks
#204755
arXiv CS
Generalized Compare-and-Swap and Space-Efficient Universal Constructions for the Infinite-Arrival Model
#204759
arXiv CS
LOSCAR-SGD: Local SGD with Communication-Computation Overlap and Delay-Corrected Sparse Model Averaging
#216798
arXiv CS
When Agents Control Robots: A Zero Trust Policy Model for Agentic Cyber-Physical Systems
#224464
arXiv CS
A Pragmatic Approach to Learned Indexing in RocksDB: Targeted Optimizations with Minimal System Modification
#224481
arXiv CS
ReMoE: Boosting Expert Reuse through Router Fine-Tuning in Memory-Constrained MoE LLM Inference
#229462
arXiv CS
When Does Deep RL Beat Calibrated Baselines? A Benchmark Study on Adaptive Resource Control
#229469
arXiv CS
PoCQ: Proof of Contribution Quality as a Lightweight Blockchain Consensus for Secure Federated Learning
#259423
arXiv CS
SIGMA: A Versatile Streaming Graph Partitioner for Vertex- and Edge-Balanced Distributed GNN Training
#259442
arXiv CS
FOLD: Fuzzy Online Deduplication for Very Large Evolving Datasets via Approximate Nearest Neighbor Search
#259448
arXiv CS
DriftSched: Adaptive QoS-Aware Scheduling under Runtime Token Drift for Multi-Tenant GPU Inference
#259449
arXiv CS
Clairvoyant: Predictive SJF Scheduling to Mitigate Head-of-Line Blocking in Serial LLM Backends
#266143
arXiv CS
Communication Strategy Selection for Multi-GPU 3D FDTD with Convolutional Perfectly Matched Boundary Layers
#266147
arXiv CS
Terastal: Layer-Variant-based Scheduling for Real-Time Multi-DNN Workloads on Heterogeneous Accelerators
#266148
arXiv CS
SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving
#267624
arXiv CS
Work Stealing for the 2D-Mesh Topology of Satellite Constellations in Low Earth Orbit
#271790
arXiv CS
From Fork-Join to Asynchronous Tasks: Parallelizing Tiled Cholesky Decomposition with OpenMP and HPX
#271809
arXiv CS
Harnessing Routing Foresight for Micro-step-level MoE load balancing in RL Post-training
#271810
arXiv CS
Beyond CPU-GPU Frequency: Memory-Clock and Tail Effects in Edge Inference Latency Estimation
#280169
arXiv CS
Evaluating Gemma4 Models as AI Teaching Assistants for Introductory Parallel Programming: A DataRaceBench Study
#280185
arXiv CS
RISE: Relay Inference and Online Scheduling for Efficient Edge-Device Collaborative Diffusion Model Services
#282769
arXiv CS
Verified Detection and Prevention of Concurrency Anomalies in Multi-Agent Large Language Model Systems
#282771
arXiv CS
Energy Efficient Scheduling of AI/ML Workloads on Multi Instance GPUs with Dynamic Repartitioning
#306998
arXiv CS
← Previous
Page 16 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.