Conceptio
›
and-cluster-computing
Topic
and-cluster-computing
Knowledge-graph topic
· documents ABOUT and-cluster-computing across the archive
656
Documents about and-cluster-computing
Documents about and-cluster-computing
Same Request, Different Answer: Quantization Amplifies Cache-Induced Divergence in LLM Serving
#971420
arXiv (All)
Benchmarking Storage Systems for Machine Learning Workloads Using NIO Bench
#972485
arXiv (All)
Beyond Lemma Sharing -- Novel Parallelization Strategies for Property Directed Reachability
#972493
arXiv (All)
DejaVu: Unifying Memory Allocations to Eliminate Redundant Copies on Unified-Memory SoCs
#972687
arXiv (All)
Sharing a Fabric with Collective Communication: Two Storage Penalties in Deep Learning Training
#974077
arXiv (All)
MicroIntent: Intent-Based Placement Strategy for Microservice Application in the Compute Continuum Using LLMs
#985334
arXiv (All)
LBFAST: A Lightweight Moment-Represented Lattice Boltzmann Solver for Multi-GPU Architectures
#986933
arXiv (All)
Avatar: Toward Autonomous End-to-End Orchestration of Scientific Workflows using LLMs
#997381
arXiv (All)
The Internet of Collaborating Things: Agentic Edge AI for Autonomous Cross-Domain Collaboration
#997548
arXiv (All)
MUC-FL: Block-Wise Marginal Utility Contribution for Communication-Efficient Federated Learning
#997551
arXiv (All)
DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management
#998060
arXiv (All)
SparseDitto: An Agentic Sparse Compilation Framework through Architecture-Aware Synthesis on GPUs
#998320
arXiv (All)
Can AI Remediate Backend Failures Safely? GuardedAct with Blast-Radius-Aware Sandboxing
#998805
arXiv (All)
Rethinking Sparse Formats for RISC-V: A Hierarchical Approach to High-Performance SpMV
#998887
arXiv (All)
Shards on a Shoestring: Empirical Characterization of NEAR Protocol Nightshade Sharding on Commodity Hardware
#1000062
arXiv (All)
Libra: Taming Attention Workload Skew in Long-Context LLM Training with Bounded Sequence Pool
#1014774
arXiv (All)
Decomposing Predictive Kubernetes Autoscaling for Large Language Model Serving Under Long Startup Delays
#1034849
arXiv (All)
Weave: Fine-Grained Dynamic SM Scheduling in an MoE Megakernel for Compute-Communication Overlap
#1035883
arXiv (All)
HyperParallel-FSDP: Topology-Aware Fully Sharded Training with Layout-Driven Muon on Ascend SuperPods
#1035990
arXiv (All)
Moser-Tardos Algorithm with small number of random bits
#163003
arXiv (OAI)
Moser-Tardos Algorithm with small number of random bits
#163962
arXiv (OAI Expanded)
Porting and Benchmarking Chapel on Emerging RISC-V Hardware: an HPC Viability Study
#609223
arXiv (OAI Expanded)
Partial FC: Training 10 Million Identities on a Single Machine
#609716
arXiv (All)
RAPID-LLM: Resilience-Aware Performance analysis of Infrastructure for Distributed LLM Training and Inference
#609852
arXiv (All)
OpenAgenet / OAN Yellow Paper: Technical Architecture for Trust-Governed Resource Identity and Discovery
#800058
arXiv (All)
A Test Taxonomy and Continuous Integration Ecosystem for Dynamic Resource Management in HPC
#820444
arXiv (All)
MineDraft: A Framework for Batch Parallel Speculative Decoding
#821432
arXiv (All)
sp-DBA: a general framework for adaptive transform-domain computation
#926788
arXiv (All)
CALM: Class-wise Agreement and Label-gated Disagreement Modulation for Decentralized Federated Learning
#973504
arXiv (All)
PACO: A Fully Cache-Oblivious Parallel FFT with One Global Redistribution
#974025
arXiv (All)
Robust Decentralized Personalized Federated Learning via Prediction-Constrained Neighborhood Collaboration
#975362
arXiv (All)
A Survey of Real-Time Support, Analysis, and Advancements in ROS 2
#987557
arXiv (All)
Tackling Parallelization Challenges of Randomized Preconditioners With Dependency Tracking
#997434
arXiv (All)
Fathom: Per-Query Read Depth for Sparse Decoding over Offloaded KV Caches
#1035613
arXiv (All)
Cloud, Edge, or Split? Profiling Onboard and Split Vision-Language Model Deployment for Drone AI
#1051908
arXiv (All)
Multi-Bin Batching for Increasing LLM Inference Throughput
#611271
arXiv (OAI Expanded)
Multi-Bin Batching for Increasing LLM Inference Throughput
#612748
arXiv (All)
Stochastic Gradient Tracking over Time-Varying Networks: One-Step Lyapunov Analysis
#618997
arXiv (OAI Expanded)
Stochastic Gradient Tracking over Time-Varying Networks: One-Step Lyapunov Analysis
#620676
arXiv (All)
Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning
#971432
arXiv (All)
Certified Split Points for Parallel Lexing: Exact and Modulo Discarded Tokens
#1013202
arXiv (All)
Verifiable Computation with Trusted Execution Environments and On-Chain Digital Rights Tokens
#1036118
arXiv (All)
Uber's Failover Architecture: Reconciling Reliability and Efficiency in Hyperscale Microservice Infrastructure
#1048587
arXiv (All)
EnFed: An Energy-aware Opportunistic Federated Learning in Resource Constrained Environments for Human Activity Recognition
#189227
arXiv (OAI)
A Biophysically-Inspired Feedback Controller for Multi-Class Cache Fairness
#609002
arXiv (OAI Expanded)
A Biophysically-Inspired Feedback Controller for Multi-Class Cache Fairness
#610064
arXiv (All)
DepTGL: A Parallel Framework for Memory-based TGNN Training with Adaptive Temporal Data Dependency Management
#619026
arXiv (OAI Expanded)
DepTGL: A Parallel Framework for Memory-based TGNN Training with Adaptive Temporal Data Dependency Management
#620705
arXiv (All)
A Cloud-Edge System for Multimodal Clinical Screening in Resource-Constrained Rural Settings
#633644
arXiv (OAI Expanded)
A Cloud-Edge System for Multimodal Clinical Screening in Resource-Constrained Rural Settings
#635993
arXiv (All)
← Previous
Page 9 of 16
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.