Conceptio
›
software-testing
Topic
software-testing
Knowledge-graph topic
· documents ABOUT software-testing across the archive
17
Documents about software-testing
Documents about software-testing
From Runnable to Shippable: Multi-Agent Test-Driven Development for Generating Full-Stack Web Applications from Requirements
#200573
arXiv CS
TRACE: A taxonomy-grounded synthetic dataset for teaching-program generation and session interpretation in Applied Behavior Analysis
#229592
arXiv CS
On the Road to Personalized Code Intelligence: Portraiting and Assisting Developers Based on Their In-IDE Behaviors
#241569
arXiv CS
Empirical Study on the Characteristics and Evolution of AI-usage in GitHub Repositories: Evidence from Code Comments
#266245
arXiv CS
Ishigaki-IDS: An Open-Weight Verifier-Aware Model for Information Delivery Specification Drafting in Building Information Modeling
#267723
arXiv CS
Do programming languages still matter to your AI coding agent teammate? Evidence at scale from chess engines
#271909
arXiv CS
No Resource, No Benchmarks, No Problem? Evaluating and Improving LLMs for Code Generation in No-Resource Languages
#282869
arXiv CS
Written by AI, Managed by AI: Semantic Space Control and Index Sickness Elimination Across 391 Consecutive Sessions
#287191
arXiv CS
Top Management Journal Portal: A Real-Source Search and Research Analytics Artifact for UTD-24 and FT50 Journals
#2673
arXiv CS
Broken Quantum: A Systematic Formal Verification Study of Security Vulnerabilities Across the Open-Source Quantum Computing Simulator Ecosystem
#2704
arXiv CS
How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks
#10440
arXiv CS
RISC-V Functional Safety for Autonomous Automotive Systems: An Analytical Framework and Research Roadmap for ML-Assisted Certification
#120601
arXiv CS
Layer-wise MoE Routing Locality under Shared-Prefix Code Generation: Token-Identity Decomposition and Compile-Equivalent Fork Redundancy
#120608
arXiv CS
Residual Risk Analysis in Benign Code: How Far Are We? A Multi-Model Semantic and Structural Similarity Approach
#126596
arXiv CS
Is this Build Failure Related to my Patch? An Empirical Study of Unrelated Build Failures in Continuous Integration
#168391
arXiv CS
Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code
#175340
arXiv CS
Debug Like a Human: Scaling LLM-based Fault Localization to Processor Design via Block-Level Instruction-Oriented Slicing
#200570
arXiv CS
No Two Developers Think Alike: How Problem-Solving Styles and Experience Shape Needs in Conversational Interaction with Copilot
#287189
arXiv CS
Measuring Curriculum Alignment across Topical Coverage, Competency, and Cognitive Depth: A Longitudinal Framework Applied to CS2013 and CS2023
#290657
arXiv CS
Developing and Testing Software for Linking Patient Data from Multiple Sources
#1852
NCBI Bookshelf
DynamicsLLM: a Dynamic Analysis-based Tool for Generating Intelligent Execution Traces Using LLMs to Detect Android Behavioural Code Smells
#10435
arXiv CS
React-ing to Grace Hopper 200: Five Open-Weights Coding Models, One React Native App, One GH200, One Weekend
#120605
arXiv CS
Operationalising Information Security Management: A Procedural Framework Analysis of ISO/IEC 27001:2022 Implementation in a Financial-Technology Organisation
#139023
arXiv CS
A Non-Destructive Methodological Framework for Modernizing Legacy Clinical Reporting Systems for AI-Driven Pharmacoinformatics: A SAS Case Study
#187387
arXiv CS
Where Do Large Language Models Fail on Competitive Programming? A Taxonomy of Failures by Algorithm Type and Difficulty Rating
#259575
arXiv CS
Lost in the Flow with Code Talkers: Unveiling the Instruction-Tuning Tax of Large Language Models in Code Tasks
#267720
arXiv CS
Are We Lost in the Woods? Detecting Silent Semantic Faults for Random Forest Classifiers with Data-informed Static Analysis
#267743
arXiv CS
Selection Without Signal, Recovery Through Expression: A Measurement Study of Post-Hoc Falsification Operators for Frozen Small Code Models
#282864
arXiv CS
Analytics for Quality Assurance for Item Pools (AQuAP): Monitoring and Maintaining Item Bank Health in AI-Driven Assessment Systems
#287195
arXiv CS
Software Testing With Large Language Models: Survey, Landscape, and Vision
#16631
Semantic Scholar
Analysis on GenAI for Source Code Scanning and Automated Software Testing
#92730
PhilArchive
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts
#155376
arXiv CS
One Developer Is All You Need: A Case Study of an AI-Augmented One-Person Squad in a Brownfield Enterprise
#200542
arXiv CS
How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions
#241567
arXiv CS
An Ocean Model Ported by a Large Language Model: Experience and Lessons from FESOM2 (Fortran to C to C++/Kokkos)
#271935
arXiv CS
From Program Slices to Causal Clarity: Evaluating Faithful, Actionable LLM-Generated Failure Explanations via Context Partitioning and LLM-as-a-Judge
#120576
arXiv CS
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
#138987
arXiv CS
The Hidden Environmental Cost of Poor Coding Practices in TensorFlow and Keras Applications: A Study on Resource Leaks and Carbon Emissions
#290649
arXiv CS
MR-Adopt: Automatic Deduction of Input Transformation Function for Metamorphic Testing
#139088
arXiv (OAI)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
#168411
arXiv CS
Software Testing Automation: Tools, Techniques, and Best Practices
#74934
PhilArchive
Concolic Testing on Individual Fairness of Neural Network Models
#189341
arXiv (OAI)
Towards Persistent Case-Based Memory for Autonomous Data Science: A CBR-Augmented R&D-Agent with a Locally Deployable Small Language Model
#259560
arXiv CS
Análisis de factibilidad de la tercerización de servicios de testing y automatización de software en Colombia
#214757
Biblioteca Digital Minerva
GakuNin RDMプロジェクトエクスポート例
#202103
Figshare
Trustworthy Self-Composable Big-Data-as-a-Service: An LLM-Orchestrated Multi-Agent Framework for Automated Data Engineering, AutoML, MLOps Deployment, and Drift-Aware Lifecycle Optimization
#282848
arXiv CS
LLM-based Contrastive Learning Approach for Bug Priority Classification
#210840
Figshare
Grammar-based test suite construction using coverage-directed algorithms over LR-graphs
#212289
SUNScholar
Toward Cybersecurity Testing and Monitoring of IoT Ecosystems
#209968
Springer Nature OA
Investigação de estratégia para redução de custo do teste de mutação com apoio de similaridade entre programas
#233454
RI UFSCar
← Previous
Page 20 of 20
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.