Topic
llms
Knowledge-graph topic · documents ABOUT llms across the archive
Documents about llms
- Teaching LLMs String Matching, Backtracking, and Error Recovery to Deduce Bases and Truth Tables for the Combinatorially Exploding Bit Manipulation Puzzles#299961arXiv CS
- Human or AI? Using Digital Behavior to Verify Essay Authorship#300183BYU ScholarsArchive
- How Far Do On-Prem Open LLMs Get on Text-to-SQL? A Cross-Family Size x Technique Frontier on BIRD#324984arXiv CS
- Capitalismo de vigilancia y grandes modelos de lenguaje#352131OSF
- A JoLT for the KV Cache: Near-Lossless KV Cache Compression via Joint Tucker and JL-Residual Allocation for LLMs#366274arXiv CS
- Roadmap for using large language models (LLMs) to accelerate cross-disciplinary research with an example from computational biology#485165Europe PMC
- Self-Bootstrapping Automated Program Repair: Using LLMs to Generate and Evaluate Synthetic Training Data for Bug Repair#611302arXiv (OAI Expanded)
- Self-Bootstrapping Automated Program Repair: Using LLMs to Generate and Evaluate Synthetic Training Data for Bug Repair#612779arXiv (All)
- Do Assessment Instruments Measure the Same Thing for Humans and LLMs? A Latent Structure Analysis#616281arXiv (OAI Expanded)
- Do Assessment Instruments Measure the Same Thing for Humans and LLMs? A Latent Structure Analysis#616781arXiv (All)
- From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection#617411arXiv (OAI Expanded)
- From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection#617911arXiv (All)
- ChatGPT, metacognitive processes and syntactic complexity : an overview of LLMs usage and its impact on writing in the EFL classroom#672488Repositorio UNAB
- When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs#674514arXiv (OAI Expanded)
- When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs#677715arXiv (All)
- Influence of structured output constraints on GPT-5-Thinking, Gemini 2.5 Pro, and open-weight LLMs for radiology protocol selection.#4139NCBI PubMed Central
- A phenomenological and formal interpretation of two experiments conducted within the cognitive environment of LLMs using the formal modelling framework of hierarchical relational ontologies#99787PhilArchive
- Navigating the ethical landscape of scholarly publishing: a comparative evaluation of Gemini and DeepSeek LLMs in addressing authorship and contributorship disputes.#125601NCBI PubMed Central
- Beyond exam accuracy: Tracking a persistent-failure set reveals visual dental reasoning gaps in multimodal LLMs.#162240PubMed
- Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!#163160arXiv (OAI)
- AI for Me, Not (Yet) for Thee? Desirable Difficulties and Deliberate Friction with LLMs#173491OSF
- A golden era for open-ended questions? Using LLMs for text classification tasks#228516OSF
- Multi-SimPsychometrican: Using LLMs Supplement Achievement Test Items based on Wright Map to Improve Measurement Precision#233233OSF
- Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning#262387IRIS
- How Should AI Talk About Us? LLMs and Social Generics#285341ODU Digital Commons
- An LLM-based agentic system for greenwashing detection#294152Digital Commons @ Michigan Tech
- Integration of Large Language Models into Computer Vision Based Traffic Monitoring and Vehicle Classification#329270Animo Repository
- Leveraging LLMs to assist tourists at destinations – The opportunities and risks of deploying open-source humanoid robots#486042SocArXiv
- Intention to Use Large Language Models Among Clinical Nurses in China With Prior Familiarity With or Experience Using LLMs: A Qualitative Study.#615636NCBI PubMed Central
- Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought#617498arXiv (OAI Expanded)
- SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs#617500arXiv (OAI Expanded)
- Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought#617998arXiv (All)
- SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs#618000arXiv (All)
- Think Inside the Chunk: RegulaRAG for Regulation-Compliant Scenario Generation using LLMs: A Case Study of UN Regulation No. 152#619109arXiv (OAI Expanded)
- Think Inside the Chunk: RegulaRAG for Regulation-Compliant Scenario Generation using LLMs: A Case Study of UN Regulation No. 152#620788arXiv (All)
- Developing an open-source framework for LLM evaluation of patients using EHR clinical documentation; performance of LLMs relative to medical professionals#629531Europe PMC
- RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs#650840arXiv (OAI Expanded)
- RefusalGuard: Geometry-Preserving Fine-Tuning for Safety in LLMs#652293arXiv (All)
- CyberQ: Generating Questions and Answers for Cybersecurity Education Using Knowledge Graph-Augmented LLMs#672048OpenAlex
- LLMs' Reshaping of People, Processes, Products, and Society in Software Development: A Comprehensive Exploration with Early Adopters#682440arXiv (OAI Expanded)
- LLMs' Reshaping of People, Processes, Products, and Society in Software Development: A Comprehensive Exploration with Early Adopters#684215arXiv (All)
- Evaluating Large Language Models for Psychometric Simulation Studies in R: Integrating Best Practices in Simulation and Prompt Design#16220OSF
- ANALYTICAL MEMORANDUM on the Capabilities of Large Language Models (LLMs) for Inferential Psychometric Profiling of Users and the Resulting Systemic Risks for Society and the State#83378PhilArchive
- Context Drift and Normative Portability: How LLMs Reshape Epistemic Authority in Institutional Reasoning - extended abstract, accepted at MBR026, Rome, June 2026, forthcoming in Springer SAPERE series#111229PhilArchive
- Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription#153033arXiv (OAI)
- LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?#158657arXiv (OAI)
- Improving Fairness on Semantic Segmentation Using Large Language Models#169459ODU Digital Commons
- Creating IT Capabilities in Workflow Automation with Unstructured Data#183492DigitalCommons@Kennesaw State University
- LLMs Struggle With Negation, but so do Humans - a Two-step Simulation Approach#196188OSF
- It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt#222583arXiv CS
Topic record · derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.