Conceptio
›
software-testing
Topic
software-testing
Knowledge-graph topic
· documents ABOUT software-testing across the archive
17
Documents about software-testing
Documents about software-testing
Compact Constraint Encoding for LLM Code Generation: An Empirical Study of Token Economics and Constraint Compliance
#2692
arXiv CS
From Translation to Superset: Benchmark-Driven Evolution of a Production AI Agent from Rust to Python
#10416
arXiv CS
OOM-RL: Out-of-Money Reinforcement Learning Market-Driven Alignment for LLM-Based Multi-Agent Systems
#10418
arXiv CS
Short Version of VERIFAI2026 Paper -- Learning Infused Formal Reasoning: Contract Synthesis, Artefact Reuse and Semantic Foundations
#13160
arXiv CS
Large-Scale Quantum Circuit Simulation on HPC Cluster via Cache Blocking, Boosting, and Gate Fusion Optimization
#13166
arXiv CS
LLM-Enhanced Log Anomaly Detection: A Comprehensive Benchmark of Large Language Models for Automated System Diagnostics
#13168
arXiv CS
Benchmarks for Trajectory Safety Evaluation and Diagnosis in OpenClaw and Codex: ATBench-Claw and ATBench-CodeX
#19077
arXiv CS
Single-Language Evidence Is Insufficient for Automated Logging: A Multilingual Benchmark and Empirical Study with LLMs
#120594
arXiv CS
How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks
#134612
arXiv CS
RAG-Reflect: Agentic Retrieval-Augmented Generation with Reflections for Comment-Driven Code Maintenance on Stack Overflow
#134623
arXiv CS
Feedback Over Form: Why Execution Feedback Matters More Than Pipeline Topology in 1-3B Code Generation
#134631
arXiv CS
Grammar-Constrained Refinement of Safety Operational Rules Using Language in the Loop: What Could Go Wrong
#139012
arXiv CS
Automating Categorization of Scientific Texts with In-Context Learning and Prompt-Chaining in Large Language Models
#139016
arXiv CS
From Threads to Trajectories: A Multi-LLM Pipeline for Community Knowledge Extraction from GitHub Issue Discussions
#141550
arXiv CS
Geographic Variation in Stack Overflow Code Quality: Evidence from a Cross-Regional Study of Coding Practices
#155363
arXiv CS
These Aren't the Reviews You're Looking For How Humans Review AI-Generated Pull Requests
#155390
arXiv CS
Ensuring Reliability in Programming Knowledge Tracing: A Re-evaluation of Attention-augmented Models and Experimental Protocols
#168403
arXiv CS
The Death Spiral of Open Source Projects: A Post-Mortem Analysis of Pull Request Workflow Dynamics
#178943
arXiv CS
User Reviews as a Source for Usability Requirements: A Precursor Study on Using Large Language Models
#180734
arXiv CS
Making OpenAPI Documentation Agent-Ready: Detecting Documentation and REST Smells with a Multi-Agent LLM System
#187381
arXiv CS
The Dangers of Non-Self-Fixed Architecture Technical Debt and Its Impact on Time-to-Fix
#195546
arXiv CS
Position: Early-Stage Quality Assurance in Annotation Pipelines Is More Cost-Effective Than Late-Stage Validation
#195550
arXiv CS
One Step Further: Understanding PLC Binaries Through Cross-Platform Reverse Engineering and Function-Level Semantic Analysis
#200567
arXiv CS
The Accessibility Capability Boundary: Operational Limits and Expansion Potential of AI-Generated Browser-Native Accessibility Systems
#204846
arXiv CS
When Web Apps Heal Themselves: A MAPE-K Based Approach to Fault Tolerance and Adaptive Recovery
#204855
arXiv CS
Beyond the Tip of the Iceberg: Understanding SATD in Dockerfiles through the Lens of Co-evolution
#216907
arXiv CS
A Tertiary Review of Large Language Model-Based Code Generating Tasks: Trends, Challenges, and Future Directions
#229579
arXiv CS
From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence
#241577
arXiv CS
Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill
#259532
arXiv CS
Code Lifespan Survival Analysis (CLSA): Predicting the Survival of Source Code Lines Using AST-Aware Mining
#259557
arXiv CS
AI-Driven Test Case Generation from Natural Language Requirements: A Survey of Techniques and Research Gaps
#266253
arXiv CS
The Windows IOCTL Census: A Corpus-Scale, Multi-Architecture Database of the Driver Control-Code Surface
#267742
arXiv CS
An Empirical Study of Gemini 3 for Detecting Natural Language Test Smells in Manual Test Cases
#271907
arXiv CS
Agents All the Way Down; A Methodology for Building Custom AI Agents from Substrate to Production
#271925
arXiv CS
Library-Aware Doubles and Iterative Repair for Large Language Model-Generated Unit Tests in OpenSIL Firmware
#290652
arXiv CS
An empirical study of LoRA-based fine-tuning of large language models for automated test case generation
#2695
arXiv CS
Participation and Power: A Case Study of Using Ecological Momentary Assessment to Engage Adolescents in Academic Research
#10414
arXiv CS
Beyond the Golden Record: Toward a Design Theory for Trustworthy Master Data Management with Self-Sovereign Identity
#10415
arXiv CS
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment
#13172
arXiv CS
SWE-TRACE: Optimizing Long-Horizon SWE Agents Through Rubric Process Reward Models and Heuristic Test-Time Scaling
#19078
arXiv CS
Persona-Based Requirements Engineering for Explainable Multi-Agent Educational Systems: A Scenario Simulator for Clinical Reasoning Training
#120606
arXiv CS
Treating Run-time Execution History as a First-Class Citizen: Co-Versioning Run-time Behavior alongside Code
#120613
arXiv CS
Finding Duplicates in 1.1M BDD Steps: cukereuse, a Paraphrase-Robust Static Detector for Cucumber and Gherkin
#124115
arXiv CS
Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework
#126595
arXiv CS
Code for All: Educational Applications of the "Vibe Coding" Hackathon in Programming Education across All Skill Levels
#134613
arXiv CS
AI-Assisted Code Review as a Scaffold for Code Quality and Self-Regulated Learning: An Experience Report
#139022
arXiv CS
Modeling Dependency-Propagated Ecosystem Impact of Changes in Maintenance Activities: Evaluating Support Strategies in the PyPI Network
#168379
arXiv CS
Zoom, Don't Wander: Why Regional Search Outperforms Pareto Reasoning and Global Optimization in Budget-Constrained SBSE
#175335
arXiv CS
Remember Your Trace: Memory-Guided Long-Horizon Agentic Framework for Consistent and Hierarchical Repository-Level Code Documentation
#187376
arXiv CS
Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
#200549
arXiv CS
← Previous
Page 19 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.