Conceptio
›
parallel-computing
Topic
parallel-computing
Knowledge-graph topic
· documents ABOUT parallel-computing across the archive
1,734
Documents about parallel-computing
Documents about parallel-computing
SwiftCache: Efficient LLM Serving for Multi-turn Conversations with Heterogeneous KV Cache Sharing
#280168
arXiv CS
S4oP: Operator-level Pruning of Structured State Space Models for Resource-Constrained Devices
#282760
arXiv CS
Urban Limits as Design Constraints: Identifying Suitable Locations for Distributed, Photovoltaic-Powered Servers
#287098
arXiv CS
SCOPE-FL: A Strategy-proof Chain-based Optimal pareto efficient Federated Learning System
#287108
arXiv CS
ARGUS: Production-Scale Tracing and Performance Diagnosis for over 10,000-GPU Clusters
#290547
arXiv CS
Asymmetry PRISM: A CPU/GPU Portfolio Optimization Engine for Deadline-Bounded Institutional Rebalancing
#299879
arXiv CS
Rethinking Burst Buffer Optimization: Enabling Layout Heterogeneity via Hybrid Analysis and LLM Guidance
#299898
arXiv CS
Bridging Design and Execution: A Visual Graph Editor for Edge and Cloud Workflows
#299905
arXiv CS
cuSBF: A Minimizer-Aware Bloom Filter for Genomic Sequence Data on Modern GPUs
#303179
arXiv CS
LMS-AR: LMS Prediction-based Adaptive Regulator for Memory Bandwidth in Multicore Systems
#303186
arXiv CS
Interference-Aware Cross-Application Placement: A Multi-Objective Optimization Approach for Microservice Clusters
#306984
arXiv CS
RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning
#310784
arXiv CS
TileMaxSim: IO-Aware GPU MaxSim Scoring with Dimension Tiling and Fused Product Quantization
#310789
arXiv CS
Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference
#321801
arXiv CS
Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving
#321803
arXiv CS
Decoupling Trust in Byzantine CRDTs: Fine-grained Post-Compromise Handling without Breaking Causality
#324861
arXiv CS
LASER: Load-Aware Serving with Early-Exit for Reasoning LLMs at the Edge
#324862
arXiv CS
Protecting Futures against Silent Data Corruption -- Efficient Task Replication for Dynamic Data Dependencies
#324869
arXiv CS
Stochastic Connectivity as the Foundation of a Runtime Model for Microservice Availability Analysis
#329056
arXiv CS
CloudyGUI: A Novel Python-based Framework for Auto-Scaling and Cloud Workload Analysis
#329058
arXiv CS
Entropy-Regularized Probabilistic Gates for Sparse Model Discovery in Scarce-Data Federated Learning
#329060
arXiv CS
Exceeding the Numerical and Performance Characteristics of IEEE-754 SGEMM with BFloat16 Tensor Cores on GPUs for Scientific Computing
#200469
arXiv CS
From the NYU Ultracomputer to Modern Exascale: A Historical and Architectural Survey of In-Network Computing and Scalable Synchronization
#280160
arXiv CS
Wattlytics: A Web Platform for Co-Optimizing Performance, Energy, and TCO in HPC Clusters
#2539
arXiv CS
Blink: CPU-Free LLM Inference by Delegating the Serving Stack to GPU and SmartNIC
#2544
arXiv CS
Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning
#2546
arXiv CS
Beyond End-to-End: Dynamic Chain Optimization for Private LLM Adaptation on the Edge
#2556
arXiv CS
SwarmIO: Towards 100 Million IOPS SSD Emulation for Next-generation GPU-centric Storage Systems
#2558
arXiv CS
MegaTrain: Full Precision Training of 100B+ Parameter Large Language Models on a Single GPU
#2571
arXiv CS
CUTEv2: Unified and Configurable Matrix Extension for Diverse CPU Architectures with Minimal Design Overhead
#10308
arXiv CS
FlexVector: A SpMM Vector Processor with Flexible VRF for GCNs on Varying-Sparsity Graphs
#10324
arXiv CS
LoDAdaC: a unified local training-based decentralized framework with adaptive gradients and compressed communication
#10325
arXiv CS
An Engineering Journey Training Large Language Models at Scale on Alps: The Apertus Experience
#13054
arXiv CS
Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents
#13065
arXiv CS
An Engineering Journey Training Large Language Models at Scale on Alps: The Apertus Experience
#14000
arXiv CS
User Experiences with MPI RMA and ULFM in a Resilient Key-Value Store Implementation
#120465
arXiv CS
e112: A Context-Aware Mobile Emergency Communication Platform Leveraging Smartphone Sensing and Cloud Services
#124012
arXiv CS
A Cloud-Native Architecture for Human-in-Control LLM-Assisted OpenSearch in Investigative Settings
#126488
arXiv CS
A comprehensive evaluation of spatial co-execution on GPUs using MPS and MIG technologies
#134522
arXiv CS
Accelerating Intra-Node GPU-to-GPU Communication Through Multi-Path Transfers with CUDA Graphs
#134524
arXiv CS
SpotVista: Availability-Aware Recommendation System for Reliable and Cost-Efficient Multi-Node Spot Instances
#138890
arXiv CS
CUDA Kernel Optimization and Counter-Free Performance Analysis for Depthwise Convolution in Cloud Environments
#141458
arXiv CS
From Barrier to Bridge: The Case for AI Data Center/Power Grid Co-Design
#155232
arXiv CS
From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications
#155234
arXiv CS
Dispatching Odyssey: Exploring Performance in Computing Clusters under Real-world Workloads
#163119
arXiv (OAI)
Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon
#175211
arXiv CS
MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production
#175220
arXiv CS
GriNNder: Breaking the Memory Capacity Wall in Full-Graph GNN Training with Storage Offloading
#178849
arXiv CS
MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters
#180623
arXiv CS
The Distributed Complexity Landscape on Trees Depends on the Knowledge About the Network Size
#180633
arXiv CS
← Previous
Page 16 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.