Conceptio
›
and-cluster-computing
Topic
and-cluster-computing
Knowledge-graph topic
· documents ABOUT and-cluster-computing across the archive
656
Documents about and-cluster-computing
Documents about and-cluster-computing
Sketching the Error, Not the Product: Post Hoc Fault Recovery for Half Precision GPU Matrix Multiplication
#1014345
arXiv (All)
Hybrid GPU-CPU Retrieval for Personalized Search at Ultra-Large Scale
#1035688
arXiv (All)
Co-occurrence Patterns of LoRA Adapters in Production Diffusion Model Inference Services
#1049064
arXiv (All)
SPECTRA: Adaptive Execution of Speculative Decoding on a Runtime-Reconfigurable Tiled Architecture
#1051246
arXiv (All)
Improved Bounds for Coin Flipping, Leader Election, and Random Selection
#179140
arXiv (OAI)
SiliconBench: Speed, Memory, and Fidelity for LLM Serving on Unified-Memory Desktops
#1013770
arXiv (All)
Competition, Collusion, and Corruption: The Spectrum of MEV Attacks on DAG-Based BFT Consensus Protocols
#1033847
arXiv (All)
Sandwich: Joint Configuration Search and Hot-Switching for Efficient CPU LLM Serving
#144330
arXiv (OAI)
Reliable Microservice Tail Latency Prediction via Decoupled Dual-Stream Learning and Gradient Modulation
#173336
arXiv (OAI)
From LLM Inference to Agentic Workloads: Characterization and Implications for Serving Systems
#611077
arXiv (OAI Expanded)
From LLM Inference to Agentic Workloads: Characterization and Implications for Serving Systems
#612554
arXiv (All)
The World's Fastest Matching Engine Algorithm
#617588
arXiv (OAI Expanded)
The World's Fastest Matching Engine Algorithm
#618088
arXiv (All)
HAPS through the Lens of Satellites and UAVs: A Function-Level Perspective on the Emerging High Altitude Economy
#623104
arXiv (OAI Expanded)
HAPS through the Lens of Satellites and UAVs: A Function-Level Perspective on the Emerging High Altitude Economy
#625377
arXiv (All)
Memory-efficient GPU pipelines for real-time non-line-of-sight reconstruction
#783347
arXiv (OAI Expanded)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
#783510
arXiv (OAI Expanded)
Memory-efficient GPU pipelines for real-time non-line-of-sight reconstruction
#784937
arXiv (All)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
#785100
arXiv (All)
Algorithmic Simplification for Million-Vertex Diffusion History Reconstruction
#789179
arXiv (OAI Expanded)
Algorithmic Simplification for Million-Vertex Diffusion History Reconstruction
#793183
arXiv (All)
Mind the Gap: The Disconnect Between Synthetic and Natural Edge Weights in Parallel Single-Source Shortest Path
#821964
arXiv (All)
The Irreversibility Budget: Fleet-Level Risk Accounting and Admission Control for Agent Operating Systems
#920427
arXiv (OAI Expanded)
The Irreversibility Budget: Fleet-Level Risk Accounting and Admission Control for Agent Operating Systems
#924661
arXiv (All)
Collaborative On-Sensor Array Cameras
#927133
arXiv (All)
Contextual Chain: A Controlled Simulation Study of Context-Informed Gossip Scheduling for Post-Partition Recovery
#974376
arXiv (All)
From Token Interfaces to Token Semantics: A Formal Composition and Conformance Model for Implementation-Neutral Token Specifications
#997553
arXiv (All)
HoliBench: A Cross-Platform Benchmarking and Deployment Toolkit for Foundation Models in CPS-IoT Applications
#1000362
arXiv (All)
Replacing Large Language Models with Jev Decision Models for Low-Latency Edge Service Orchestration
#1037436
arXiv (All)
XIR: A Framework for Interoperability across Cross-Chain Protocols Based on a Verifiable Intermediate Representation
#1048803
arXiv (All)
Cache Your Prompt When It's Green: Carbon-Aware Caching for Large Language Model Serving
#139171
arXiv (OAI)
PAS-QFL: Personalized Ansatz Selection for Quantum Federated Learning under Client Data Heterogeneity
#610954
arXiv (OAI Expanded)
PAS-QFL: Personalized Ansatz Selection for Quantum Federated Learning under Client Data Heterogeneity
#612431
arXiv (All)
Thermo-FL: Thermal-Aware Robust Federated Fine-Tuning of Large Language Models for Edge AI
#661450
arXiv (OAI Expanded)
Thermo-FL: Thermal-Aware Robust Federated Fine-Tuning of Large Language Models for Edge AI
#662500
arXiv (All)
Agentic-Kube: A Graph-Enhanced Multi-Agent Reinforcement Learning Framework for Multi-Objective Kubernetes Scheduling
#682607
arXiv (OAI Expanded)
Agentic-Kube: A Graph-Enhanced Multi-Agent Reinforcement Learning Framework for Multi-Objective Kubernetes Scheduling
#684382
arXiv (All)
Great Expectations: Benchmarking the Real-World Performance of RVV 1.0 in HPC
#783267
arXiv (OAI Expanded)
Great Expectations: Benchmarking the Real-World Performance of RVV 1.0 in HPC
#784857
arXiv (All)
FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effective LLM Inference
#808979
arXiv (OAI Expanded)
Audit-First Rollback Semantics for Safety-Critical Deployment Pipelines
#809858
arXiv (OAI Expanded)
FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effective LLM Inference
#818971
arXiv (All)
Audit-First Rollback Semantics for Safety-Critical Deployment Pipelines
#819854
arXiv (All)
MeanField Surrogate Modeling for Scalable Runtime Scheduling of Concurrent Heterogeneous AI Inference on Shared GPUs
#822511
arXiv (All)
Toward a Formally Verified Optimality Certificate for OGR(29): A SAT-Encoding Methods Note with Small-Case Demos
#972488
arXiv (All)
Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training
#975168
arXiv (All)
Building py-kvcache: A Performance Characterization of External KV Caching for vLLM with NVMe SSDs
#999433
arXiv (All)
ForgeMegakernel: A General Framework for Efficient Auto-Regressive Model Decode Megakernels
#1000332
arXiv (All)
SSD-LLaMA: SSD-Native Inference for Trillion-Parameter MoE at 1+ Token/s on a Consumer PC
#1011809
arXiv (All)
Towards Training Private LLMs: Exploring Fine-Tuning Language Models on Apple Silicon with RDMA over Thunderbolt
#1034214
arXiv (All)
← Previous
Page 12 of 16
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.