Conceptio
›
software-testing
Topic
software-testing
Knowledge-graph topic
· documents ABOUT software-testing across the archive
17
Documents about software-testing
Documents about software-testing
Beyond Single Reports: Evaluating Automated ATT&CK Technique Extraction in Multi-Report Campaign Settings
#2689
arXiv CS
ReCodeAgent: A Multi-Agent Workflow for Language-agnostic Translation and Validation of Large-scale Repositories
#2690
arXiv CS
An Empirical Analysis of Static Analysis Methods for Detection and Mitigation of Code Library Hallucinations
#6041
arXiv CS
E2E-REME: Towards End-to-End Microservices Auto-Remediation via Experience-Simulation Reinforcement Fine-Tuning
#10425
arXiv CS
Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety
#13164
arXiv CS
Learning Project-wise Subsequent Code Edits via Interleaving Neural-based Induction and Tool-based Deduction
#13167
arXiv CS
How Developers Adopt, Use, and Evolve CI/CD Caching: An Empirical Study on GitHub Actions
#14104
arXiv CS
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
#19085
arXiv CS
Bridging the Gap between User Intent and LLM: A Requirement Alignment Approach for Code Generation
#31334
arXiv CS
Project Prometheus: Bridging the Intent Gap in Agentic Program Repair via Reverse-Engineered Executable Specifications
#120598
arXiv CS
WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning
#124118
arXiv CS
CrossCommitVuln-Bench: A Dataset of Multi-Commit Python Vulnerabilities Invisible to Per-Commit Static Analysis
#126577
arXiv CS
Institutionalizing Best Practices in Research Computing: A Framework and Case Study for Improving User Onboarding
#126579
arXiv CS
VLAA-GUI: Knowing When to Stop, Recover, and Search, A Modular Framework for GUI Automation
#126588
arXiv CS
A Comparison of ROS 2 and AUTOSAR Adaptive Platform Against Industry-Elicited Automotive Middleware Requirements
#134617
arXiv CS
A Systematic AI Adoption Framework for Higher Education: From Student GenAI Usage to Institutional Integration
#134627
arXiv CS
On the Footprints of Reviewer Bots Feedback on Agentic Pull Requests in OSS GitHub Repositories
#138993
arXiv CS
Toward a Science of Intent: Closure Gaps and Delegation Envelopes for Open-World AI Agents
#141565
arXiv CS
Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination
#141567
arXiv CS
From TinyGo to gc Compiler: Extending Zorya's Concolic Framework to Real-World Go Binaries
#155367
arXiv CS
ARISE: A Repository-level Graph Representation and Toolset for Agentic Fault Localization and Program Repair
#155378
arXiv CS
Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection
#155381
arXiv CS
HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems
#155382
arXiv CS
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
#155383
arXiv CS
ARIADNE: Agentic Reward-Informed Adaptive Decision Exploration via Blackboard-Driven MCTS for Competitive Program Generation
#155386
arXiv CS
A meta-analysis of the effect of generative AI on productivity and learning in programming
#168399
arXiv CS
AISSA: Implementation and Deployment of an AI-based Student Slides Analysis tool for Academic Presentations
#168402
arXiv CS
CodeEvolve: LLM-Driven Evolutionary Optimization with Runtime-Enriched Target Selection for Multi-Language Code Enhancement
#168406
arXiv CS
Instruction Adherence in Coding Agent Configuration Files: A Factorial Study of Four File-Structure Variables
#175327
arXiv CS
Property-Level Reconstructability of Agent Decisions: An Anchor-Level Pilot Across Vendor SDK Adapter Regimes
#178941
arXiv CS
Breaking the Dependency Chaos: A Constraint-Driven Python Dependency Resolution Strategy with Selective LLM Imputation
#178945
arXiv CS
Divergent Multi-Version Execution (DME): Canonical Instruction-Trace Fault Detection via Structural Address-Space Decorrelation
#180736
arXiv CS
Is Agentic AI Ready for Real-World Hardware Engineering? A Deep Dive with Phoenix-bench
#195562
arXiv CS
Three Heads Are Better Than One: A Multi-perspective Reasoning Framework for Enhanced Vulnerability Detection
#200545
arXiv CS
AgentModernize: Preserving Business Logic in Legacy Modernization with Multi-Agent LLMs and Behavioral Specification Graphs
#200560
arXiv CS
AI Policy, Disclosure, and Human in the Loop: How Are Contribution Guidelines Adapting to GenAI?
#200579
arXiv CS
What's Inside a GitHub Repository? An Empirical Study on the Contents of 10K Projects
#200580
arXiv CS
TARIPlay: A Test Framework for AR Applications based on Interactive Area Tracking in Playback Videos
#200583
arXiv CS
Library Drift: Diagnosing and Fixing a Silent Failure Mode in Self-Evolving LLM Skill Libraries
#204848
arXiv CS
SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations
#216882
arXiv CS
Deterministic vs. Probabilistic Summarisation: An Empirical Trade-off Study in Design Pattern Centric Java Code
#216892
arXiv CS
Stdlib or Third-Party? Empirical Performance and Correctness of LLM-Assisted Zero-Dependency Python Libraries
#216900
arXiv CS
Uncovering multi-channel magnetic hopfion annihilation via a single-node, billion-spin-scale atomistic framework
#229574
arXiv CS
Revisiting Vul-RAG: Reproducibility and Replicability of RAG-based Vulnerability Detection with Open-Weight Models
#259564
arXiv CS
TICoder: A Repository-Level Code Generation Framework with Test-Driven Planning and Implementation-Aware Reuse
#267730
arXiv CS
Understanding the Rejection of Fixes Generated by Agentic Pull Requests -- Insights from the AIDev Dataset
#271910
arXiv CS
Beyond Problem Solving: UOJ-Bench for Evaluating Code Generation, Hacking, and Repair in Competitive Programming
#271916
arXiv CS
The Perils of Agency: How Developers Perceive, Prioritize, and Address Risks in Agentic AI Products
#282885
arXiv CS
Prompt Quality and Pull Request Outcomes: A Stage-Based Empirical Study of LLM-Assisted Development
#290653
arXiv CS
Is US Defense Acquisition Ready to Acquire AI-Enabled Capabilities? Assessing the DoD Software Acquisition Pathway Through a Scenario-Based Policy Analysis
#266237
arXiv CS
← Previous
Page 18 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.