Conceptio
›
cloud-computing
Topic
cloud-computing
Knowledge-graph topic
· documents ABOUT cloud-computing across the archive
170
Documents about cloud-computing
Documents about cloud-computing
Tropical: Enhancing SLO Attainment in Disaggregated LLM Serving via SLO-Aware Multiplexing
#280165
arXiv CS
Raiders of the Lost Log: Synchronous Parallel In-Place Models and Algorithms
#280170
arXiv CS
A RAG-Enhanced Bi-Level Cognitive Orchestration Framework for LEO Satellite Networks
#280181
arXiv CS
From GPU to Microcontroller: Online Ridge Regression for Edge-Deployable Traffic Prediction
#282765
arXiv CS
Spotlight: Synergizing Seed Exploration and Spot GPUs for DiT RL Post-Training
#287095
arXiv CS
The Energy Consumption of Transformer Fine-Tuning: A Roofline-Inspired Scaling Model
#299876
arXiv CS
When Staking Rewards Compound: Measuring the Impact of Ethereum's Pectra Upgrade
#299880
arXiv CS
Efficient Network Inference via Hardware-Aware Architecture Search, Model Pruning & Quantization
#299883
arXiv CS
Clutch: High Performance Vector-Scalar Comparison using DRAM via Chunked Temporal Coding
#299888
arXiv CS
StickyInvoc: Rethinking Task Models for High-throughput Workflows in the LLM Era
#299897
arXiv CS
KineticSim: A Lightweight, High-Performance Execution Engine for Real-Time Market Simulators
#299900
arXiv CS
Semantic Lock: Synchronization Based on the Analysis of the Operation Conflict Graph
#303182
arXiv CS
An Efficient Construction of Completely Independent Spanning Trees in Dense Gaussian Networks
#303187
arXiv CS
NEURON-Fabric: Architecture-Runtime Co-Design for Controlled Low-Bit Gradient Communication
#306986
arXiv CS
Endeavor: Efficient PairHMM for Detection of DNA Variants in Genome-Scale Datasets
#306987
arXiv CS
CV-Rules: Serializability Verification of Concurrency Control Protocols via Explicit Transaction Ordering
#306993
arXiv CS
Power-Flexible AI Data Centers: A New Paradigm for Grid-Responsive Compute
#306996
arXiv CS
CHAMB-GA: A Containerized HPC Scalable Microservice-Based Framework for Genetic Algorithms
#310781
arXiv CS
Priceless: An examination of Serverless Functions-as-a-Service (FaaS) pricing models
#310790
arXiv CS
YAIFS: Yet (not) Another Intelligent Fog Simulator: A Framework for Agent-Driven Computing Continuum Modeling & Simulation
#124028
arXiv CS
Parallel Schwarz Alternating Methods Implemented on Cloud Architecture for Solving the Coupled Problem of Electrophoresis
#239154
HAL (France)
NL-CPS: Reinforcement Learning-Based Kubernetes Control Plane Placement in Multi-Region Clusters
#2535
arXiv CS
Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS
#2547
arXiv CS
Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start
#2559
arXiv CS
ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache
#2562
arXiv CS
Fine-Grained Power and Energy Attribution on AMD GPU/APU-Based Exascale Nodes
#2563
arXiv CS
DeepStack: Scalable and Accurate Design Space Exploration for Distributed 3D-Stacked AI Accelerators
#2572
arXiv CS
MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs
#5935
arXiv CS
Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures
#5938
arXiv CS
Sensor Placement for Tsunami Early Warning via Large-Scale Bayesian Optimal Experimental Design
#5940
arXiv CS
Fine-Grained Power and Energy Attribution on AMD GPU/APU-Based Exascale Nodes
#5941
arXiv CS
Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter
#18981
arXiv CS
CoCoDiff: Optimizing Collective Communications for Distributed Diffusion Transformer Inference Under Ulysses Sequence Parallelism
#18988
arXiv CS
Trust, but Verify: ByzTwin-Range, a Digital Twin Cyber-Range for Byzantine Faults
#120466
arXiv CS
Predictive Sectorization and Bayesian Optimized Consensus for Admission Control in Autonomous Airspace Operations
#120484
arXiv CS
Stream-CQSA: Avoiding Out-of-Memory in Attention Computation via Flexible Workload Scheduling
#124007
arXiv CS
FEPLB: Exploiting Copy Engines for Nearly Free MoE Load Balancing in Distributed Training
#124018
arXiv CS
Distributed Generative Inference of LLM at Internet Scales with Multi-Dimensional Communication Optimization
#126492
arXiv CS
FlashSpread: IO-Aware GPU Simulation of Non-Markovian Epidemic Dynamics via Kernel Fusion
#134527
arXiv CS
TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training
#138895
arXiv CS
SDSL-Solver: Scalable Distributed Sparse Linear Solvers for Large-Scale Interior Point Methods
#138899
arXiv CS
Spark Policy Toolkit: Semantic Contracts and Scalable Execution for Policy Learning in Spark
#141461
arXiv CS
PolyKV: A Shared Asymmetrically-Compressed KV Cache Pool for Multi-Agent LLM Inference
#141462
arXiv CS
FedPLT: Scalable, Resource-Efficient, and Heterogeneity-Aware Federated Learning via Partial Layer Training
#155239
arXiv CS
Heterogeneous Model Fusion for Privacy-Aware Multi-Camera Surveillance via Synthetic Domain Adaptation
#155245
arXiv CS
nvPAX: Constrained Optimization for Dynamic Power Allocation in Hierarchical and Multi-Tenant Systems
#155252
arXiv CS
Hierarchical Federated Learning for Networked AI: From Communication Saving to Architecture-Aware Design
#155266
arXiv CS
Relay Buffer Independent Communication over Pooled HBM for Efficient MoE Inference on Ascend
#168270
arXiv CS
From Coordinate Matching to Structural Alignment: Rethinking Prototype Alignment in Heterogeneous Federated Learning
#168271
arXiv CS
Multi-Tier Labeling and Physics-Informed Learning for Orbital Anomaly Detection at Scale
#175207
arXiv CS
← Previous
Page 14 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.