Conceptio
›
software-testing
Topic
software-testing
Knowledge-graph topic
· documents ABOUT software-testing across the archive
17
Documents about software-testing
Documents about software-testing
Key Developer Roles and Organizational Coupling in Microservices: A Longitudinal Analysis
#141553
arXiv CS
R$^3$-SQL: Ranking Reward and Resampling for Text-to-SQL
#141561
arXiv CS
BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks
#141568
arXiv CS
KVerus: Scalable and Resilient Formal Verification Proof Generation for Rust Code
#155360
arXiv CS
TeamUp: Semantic Project Matching and Team Formation for Learning at Scale
#155373
arXiv CS
Learning Correct Behavior from Examples: Validating Sequential Execution in Autonomous Agents
#155377
arXiv CS
A Low-Code Approach for the Automatic Personalization of Conversational Agents
#155389
arXiv CS
HEJ-Robust: A Robustness Benchmark for LLM-Based Automated Program Repair
#155393
arXiv CS
Beyond Translation Accuracy: Addressing False Failures in LLM-Based Code Translation
#155394
arXiv CS
Constraint Decay: The Fragility of LLM Agents in Backend Code Generation
#168374
arXiv CS
Exploring the Effectiveness of Abstract Syntax Tree Patterns for Algorithm Recognition
#168384
arXiv CS
From Chat to Interview: Agentic Requirements Elicitation with an Experience Ontology
#168388
arXiv CS
Conflict Essences for Transformation Rules with Nested Application Conditions -- Long Version
#168393
arXiv CS
SynConfRoute: Syntax-Aware Routing for Efficient Code Completion with Small CodeLLMs
#168394
arXiv CS
Toward an Understanding of Developer Behaviour while Using Bug Localization Tools
#168398
arXiv CS
Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
#168412
arXiv CS
CppPerf: An Automated Pipeline and Dataset for Performance-Improving C++ Commits
#175314
arXiv CS
Scaling Qubit Mapping and Routing With Position Graph Abstraction and Memoization
#175342
arXiv CS
Using Semantic Distance to Estimate Uncertainty in LLM-Based Code Generation
#175348
arXiv CS
EvidenT: An Evidence-Preserving Framework for Iterative System-Level Package Repair
#175353
arXiv CS
StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning
#178942
arXiv CS
An Extensive Replication Study of the ABLoTS Approach for Bug Localization
#178944
arXiv CS
SHIA: A Direct SysML-Hardware Interface Architecture for Model-Centric Verification
#178949
arXiv CS
SieveFL: Hierarchical Runtime-Aware Pruning for Scalable LLM-Based Fault Localization
#180719
arXiv CS
Improving Code Translation with Syntax-Guided and Semantic-aware Preference Optimization
#180724
arXiv CS
SWE-Cycle: Benchmarking Code Agents across the Complete Issue Resolution Cycle
#180726
arXiv CS
SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades
#187379
arXiv CS
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
#195557
arXiv CS
PROTEA: Offline Evaluation and Iterative Refinement for Multi-Agent LLM Workflows
#200547
arXiv CS
NOETHER: A Constructive Framework for Metamorphic Pattern Discovery from Operator Algebras
#200568
arXiv CS
Low-Code Paradox in DevOps: Security and Governance Insights from Practitioners
#200577
arXiv CS
Can LLMs Produce Better Object-Oriented Designs than Human-Involved Development?
#204840
arXiv CS
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
#204845
arXiv CS
Why Are Agentic Pull Requests Merged or Rejected? An Empirical Study
#216883
arXiv CS
Quality and Security Signals in AI-Generated Python Refactoring Pull Requests
#216898
arXiv CS
RefusalBench: Why Refusal Rate Misranks Frontier LLMs on Biological Research Prompts
#216910
arXiv CS
Articulate but Wrong: Self-Review Failures in LLM-Based Code Modernization
#216913
arXiv CS
HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML
#229564
arXiv CS
A Heuristic Approach to Localize CSS Properties for Responsive Layout Failures
#229580
arXiv CS
Temporal Modeling of Change History for Black-Box Test Suite Minimization
#229582
arXiv CS
Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction
#229594
arXiv CS
SmellDoc: Extending Elastic Stack for Microservice Bad Smell Detection and Visualization
#229598
arXiv CS
Usability Analysis of Configurator User Interfaces with Multimodal Large Language Models
#241566
arXiv CS
Beyond pass@k: Redundancy-Aware RLVR for Multi-Sample Code Generation
#241582
arXiv CS
Towards the Readability of LLM-Generated Codes through Multitask Representation Engineering
#259536
arXiv CS
SmellBench: Towards Fine-Grained Evaluation of Code Agents on Refactoring Tasks
#259548
arXiv CS
ADK Arena: Evaluating Agent Development Kits via LLM-as-a-Developer
#259549
arXiv CS
Toward a Generalized Defense Across Sparse, Continuous, and Structured Parameter Attacks
#259572
arXiv CS
Agentic Very Much! Adoption of Coding Agent in New GitHub Projects
#266236
arXiv CS
Declarative Skills for AI Agents in Knowledge-Grounded Tool-Use Workflows
#266243
arXiv CS
← Previous
Page 12 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.