Conceptio
›
parallel-computing
Topic
parallel-computing
Knowledge-graph topic
· documents ABOUT parallel-computing across the archive
1,734
Documents about parallel-computing
Documents about parallel-computing
An Empirical Evaluation of Quantum-Inspired QUBO Methods for Heterogeneous HPC Workflow Mapping and Scheduling
#224468
arXiv CS
DECICE: AI-Driven Scheduling and Digital Twin Integration for the Cloud-HPC-Edge Compute Continuum
#224471
arXiv CS
A Methodology to Assess Power Modeling in Energy-Aware Federated Learning on Heterogeneous Mobile Devices
#238585
arXiv CS
EES-CND: Collaborative Neural Decision-Making for Drift-Aware Fault-Tolerant Edge-Cloud Service Placement
#246505
arXiv CS
Compliance-Scored Best-of-N Guardrail Orchestration for Multimodal Document Generation in Payments Dispute Defense
#246517
arXiv CS
The Cartan-Topos Protocol: A Unified Geometric and Categorical Framework for Resilient Multi-Agent Coordination
#246530
arXiv CS
EvalStop: Using World Feedback to Detect and Correct Reward Overoptimization in Multi-Tenant RLHF Platforms
#259436
arXiv CS
UltraEP: Unleash MoE Training and Inference on Rack-Scale Nodes with Near-Optimal Load Balancing
#259437
arXiv CS
BlobShuffle: Cost-Effective Repartitioning in Stream Processing Systems via Object Storage Exemplified with Kafka Streams
#259445
arXiv CS
Aperon Technical Report: Hierarchical No-Pointer Tangent-Local Search for High-Dimensional Approximate Nearest Neighbors
#267622
arXiv CS
AlignFed: Alignment-Aware Asynchronous Federated Fine-Tuning for Large Language Models in Heterogeneous Edge Environments
#267628
arXiv CS
Extreme-Scale Atomistic Simulation of Real-Temperature Magnetic Skyrmion Dynamics by Coupled Spin-Lattice Modeling
#271780
arXiv CS
Compressed-Resident Genomics: Full-Pipeline Device-Resident GPU LZ77 Decode with Position-Invariant Random Access
#287099
arXiv CS
The Sheaf Laplacian: A Topological Framework for Data Fusion and Consensus in Distributed Sensing Networks
#290556
arXiv CS
Node-Level Performance and Energy Characterization of Flagship Science Applications on SuperMUC-NG Phase 2
#299881
arXiv CS
CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregation
#303178
arXiv CS
OpenMP GPU Acceleration and Portability of TRIMEG-C1 for Electromagnetic Gyrokinetic Simulations in Tokamak Plasmas
#303181
arXiv CS
FBench: A Flexible Benchmark for CFG-Based What-If Exploration of HPC I/O Patterns
#321799
arXiv CS
Five Ways to Build a Concurrent Linked From Coarse-Grain Locking to Lock-Free Algorithms
#321812
arXiv CS
CLEAR-MoE: Shared-Basis Expert Extraction from Frozen Vision Transformers via Calibration-Driven Layer Selection
#321816
arXiv CS
Omni-Flow: A Unified Workflow Orchestration and Distributed KV Cache Sharing Framework for Multimodal Inference
#324866
arXiv CS
On the Decidability of Distributed Tasks with Output Sets under Asynchrony and Any Number of Crashes
#2552
arXiv CS
Tight Bounds on Window Size and Time for Single-Agent Graph Exploration under T-Interval Connectivity
#2578
arXiv CS
T-RBFT: A Scalable and Efficient Byzantine Consensus Based on Trusted Execution Environment for Consortium Blockchain
#31241
arXiv CS
Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM
#124032
arXiv CS
Guess-Verify-Refine: Data-Aware Top-K for Sparse-Attention Decoding on Blackwell via Temporal Correlation
#134523
arXiv CS
Unfolding an Atomistic World: Atomistic Simulation of Reactor Pressure Vessel Steel Across Year-and-Meter Scales
#138894
arXiv CS
A Workflow-Oriented Framework for Asynchronous Human-AI Collaboration in Hybrid and Compute-Intensive HPC Environments
#155226
arXiv CS
VUDA: Breaking CUDA-Vulkan Isolation for Spatial Sharing of Compute and Graphics on the Same GPU
#155257
arXiv CS
The Illusion of Power Capping in LLM Decode: A Phase-Aware Energy Characterisation Across Attention Architectures
#178843
arXiv CS
Accelerating State-Vector Quantum Simulation on Integrated GPUs via Cache Locality Optimization: A Cross-Architecture Evaluation
#187278
arXiv CS
Mat2Boundary: Treating User-Defined Boundary Condition as SpMV for Distributed PDE Solvers on Block-Structured Grids
#187281
arXiv CS
DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration
#187285
arXiv CS
From Backup Restoration to Minimum Viable Factory Recovery: A Systematization of Ransomware Recovery in Manufacturing Systems
#194282
arXiv CS
CB-SpMV:A Data Aggregating and Balance Algorithm for Cache-Friendly Block-Based SpMV on GPUs
#200446
arXiv CS
Rotary GPU: Exploring Local Execution Paths for Large Mixture-of-Experts Models Under Limited GPU Memory
#238571
arXiv CS
The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution
#238586
arXiv CS
Not All Errors Are Equal: A Systematic Study of Error Propagation in Large Language Model Inference
#246503
arXiv CS
Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode
#246541
arXiv CS
Demystifying NVSHMEM: A System-Level Analysis on Symmetric Memory and Device-Initiated Operations in GPU Communication
#259420
arXiv CS
The Usefulness Gap in Proof-of-Useful-Work: An Empirical Study of Pearl's cuPOW Protocol
#259428
arXiv CS
P-Cast Precision in FP8 Attention: Sink-Induced Collapse and the Optimality of S=2^8
#266150
arXiv CS
A Low-Latency Semantic State Estimator using Latent Predictive Learning for Dynamic Network Monitoring and Orchestration
#267620
arXiv CS
When Errors Become Narratives: A Longitudinal Taxonomy of Silent Failures in a Production LLM Agent Runtime
#271774
arXiv CS
PLAIground: SLO-Driven Runtime Model Selection for Compound AI Systems in the Edge-Cloud-Space Continuum
#271777
arXiv CS
SupraSNN: Exploiting Synapse-Level Parallelism in Spiking Neural Network Accelerators through Co-Optimized Mapping and Scheduling
#271789
arXiv CS
SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture
#271797
arXiv CS
Rigel: Reverse-Engineering the Metal 4.1 Tensor Compute Path on the Apple M4 Max GPU
#271799
arXiv CS
Consensus Time in 3-Majority and 2-Choices Is Determined by the Maximum Initial Opinion Density
#271812
arXiv CS
When the Next Step Is Not One Step: Distribution-Aware Execution Modeling for Concurrent Go Programs
#282768
arXiv CS
← Previous
Page 18 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.