Conceptio
›
software-testing
Topic
software-testing
Knowledge-graph topic
· documents ABOUT software-testing across the archive
17
Documents about software-testing
Documents about software-testing
Mono2Sls: Automated Monolith-to-Serverless Migration via Multi-Stage Pipeline with Static Analysis
#138990
arXiv CS
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents
#139009
arXiv CS
AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking
#139011
arXiv CS
CUJBench: Benchmarking LLM-Agent on Cross-Modal Failure Diagnosis from Browser to Backend
#139015
arXiv CS
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
#139018
arXiv CS
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
#155375
arXiv CS
From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work
#168375
arXiv CS
On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
#168387
arXiv CS
Operationalizing Ethics for AI Agents: How Developers Encode Values into Repository Context Files
#168390
arXiv CS
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
#175320
arXiv CS
Correct-by-Construction G-Code Generation: A Neuro-Symbolic Approach via Separation Logic
#175321
arXiv CS
SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications
#175336
arXiv CS
ParityFuzz: Finding Inconsistencies across Solidity Compilers via Fine-Grained Mutation and Differential Analysis
#175346
arXiv CS
Method-level Change-proneness: A Better Metric for Black-box Test Suite Minimization
#195560
arXiv CS
Automated Root-Cause Subclassification and No-Code Fix Generation for Invalid Bug Reports
#200557
arXiv CS
SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering
#200561
arXiv CS
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation
#200574
arXiv CS
"Refactoring Runaway": Understanding and Mitigating Tangled Refactorings in Coding Agents for Issue Resolution
#216884
arXiv CS
PITMuS: A Tool for Automated Bug Dataset Generation via Source-Level Mutant Reconstruction
#216893
arXiv CS
Evidence Absence Is Not Evidence Insufficiency: Diagnosing NEI Construction Artifacts in Fact Verification
#229566
arXiv CS
Names Are All You Need: Effective and Safe Regression Test Selection for Python
#229585
arXiv CS
Agora: Toward Autonomous Bug Detection in Production-Level Consensus Protocols with LLM Agents
#241560
arXiv CS
Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA
#241570
arXiv CS
LLM Based Web Accessibility Repair: An Empirical Study of Detection, Remediation, and Cost
#241586
arXiv CS
OptiLoop: Coordination-in-the-Loop Verification and Repair for LLM-Generated Optimization Agents
#241587
arXiv CS
Asuka-Bench: Benchmarking Code Agents on Underspecified User Intent and Multi-Round Refinement
#259539
arXiv CS
Description-Code Inconsistency in Real-world MCP Servers: Measurement, Detection, and Security Implications
#259563
arXiv CS
A Causal Probabilistic Framework for Perception-Informed Closed-Loop Simulation of Autonomous Driving
#266240
arXiv CS
When LLMs Invent Rust Crates: An Empirical Study of Hallucination Patterns and Mitigation
#267727
arXiv CS
Mining Architectural Quality Under Agentic AI Adoption: A Causal Study of Java Repositories
#271912
arXiv CS
Finding Compiler-Platform Interaction Bugs in Deep Learning Pipelines via Cross-Layer Constraints
#287198
arXiv CS
Beyond Static Endpoints: Tool Programs as an Interface for Flexible Agentic Web Services
#290645
arXiv CS
Source-Free Detection and Impact Analysis of Compiler Optimization Problems in Mobile Applications
#299991
arXiv CS
Ensuring Open Source Integrity: The Intersection of Copy-Based Reuse and License Compliance
#299992
arXiv CS
Generate with CodeXHug: A Dataset to Enhance Model Cards with Code Usage Patterns
#299994
arXiv CS
A Set-Theoretic Approach to Detecting Logic Bugs in DBMS Inner Join Optimizations
#299995
arXiv CS
FGDM: Reasoning Aware Multi-Agentic Framework for Software Bug Detection using Chain of Thought and Tree of Thought Prompting
#141570
arXiv CS
Human oversight of agentic systems in practice: Examining the oversight work, challenges, and heuristics of developers using software agents
#259552
arXiv CS
The State of Peer Review in Empirical Software Engineering: A Community Survey on Review Load, Quality, and GenAI Use
#259565
arXiv CS
Design and Development of MATLAB-based software for running a 3-point bending fatigue testing apparatus
#208568
Figshare
Externalization in LLM Agents: A Unified Review of Memory, Skills, Protocols and Harness Engineering
#2665
arXiv CS
An Empirical Study on Influence-Based Pretraining Data Selection for Code Large Language Models
#2681
arXiv CS
From OSS to Open Source AI: an Exploratory Study of Collaborative Development Paradigm Divergence
#6033
arXiv CS
An Eye for Trust: An Exploration of Developers' Trust Perceptions Through Urgency and Reputation
#6037
arXiv CS
Automated BPMN Model Generation from Textual Process Descriptions: A Multi-Stage LLM-Driven Approach
#13174
arXiv CS
Structured Safety Auditing for Balancing Code Correctness and Content Safety in LLM-Generated Code
#13175
arXiv CS
Vibe-Coding: Feedback-Based Automated Verification with no Human Code Inspection, a Feasibility Study
#19076
arXiv CS
Towards Better Static Code Analysis Reports: Sentence Transformer-based Filtering of Non-Actionable Alerts
#120572
arXiv CS
Cache-Related Smells in GitLab CI/CD: Comprehensive Catalog, Automated Detection, and Empirical Evidence
#120585
arXiv CS
From Language to Action: Enhancing LLM Task Efficiency with Task-Aware MCP Server Recommendation
#120604
arXiv CS
← Previous
Page 16 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.