Conceptio
›
computation-and-language
Topic
computation-and-language
Knowledge-graph topic
· documents ABOUT computation-and-language across the archive
4,635
Documents about computation-and-language
Documents about computation-and-language
SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA
#629835
arXiv (All)
Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent
#630101
arXiv (All)
LENS: In-Context Search via Latent Evidence Exploration over Dynamic Raw Documents
#630171
arXiv (All)
Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents
#633385
arXiv (OAI Expanded)
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
#633445
arXiv (OAI Expanded)
Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents
#635734
arXiv (All)
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
#635794
arXiv (All)
Mitigating Identity Essentialism in LLM Agents with Longitudinal Life Trajectories
#656772
arXiv (OAI Expanded)
AgentMercury: Your Agent Can Synthesize Verifiable Environments for Business Scenarios at scale
#657060
arXiv (OAI Expanded)
Authorship without writing: large language models and the senior author analogy
#232869
Ora Oxford University Research
SocialCoach: Personalized Social Skill Learning with Agentic Tutoring and Practice
#614082
arXiv (OAI Expanded)
SocialCoach: Personalized Social Skill Learning with Agentic Tutoring and Practice
#615049
arXiv (All)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
#611240
arXiv (OAI Expanded)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
#612717
arXiv (All)
Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test
#128412
arXiv (OAI)
Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data
#141713
arXiv (OAI)
One RL to See Them All: Visual Triple Unified Reinforcement Learning
#147093
arXiv (OAI)
Creating and Evaluating Personas Using Generative AI: A Scoping Review of 81 Articles
#149250
arXiv (OAI)
#PraCegoVer: A Large Dataset for Image Captioning in Portuguese
#152947
arXiv (OAI)
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding
#153071
arXiv (OAI)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
#153108
arXiv (OAI)
Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training
#158758
arXiv (OAI)
VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding
#609151
arXiv (OAI Expanded)
DYNASHIELD: A Black-Box Moving Target Defense for LLMs via Dynamic Decoding Customization
#611272
arXiv (OAI Expanded)
Bye-bye, Bluebook? Automating Legal Drudgery With AI-Augmented Rule Following
#611300
arXiv (OAI Expanded)
Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition
#611369
arXiv (OAI Expanded)
DYNASHIELD: A Black-Box Moving Target Defense for LLMs via Dynamic Decoding Customization
#612749
arXiv (All)
Bye-bye, Bluebook? Automating Legal Drudgery With AI-Augmented Rule Following
#612777
arXiv (All)
Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition
#612846
arXiv (All)
VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?
#614213
arXiv (OAI Expanded)
When AI Rewrites, Classifiers Relax: Uncertainty-Aware Sentiment Analysis on Sarcastic and AI-Paraphrased Social Text
#614341
arXiv (OAI Expanded)
VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?
#615180
arXiv (All)
When AI Rewrites, Classifiers Relax: Uncertainty-Aware Sentiment Analysis on Sarcastic and AI-Paraphrased Social Text
#615308
arXiv (All)
Whose Gold? Annotator-Pool Disagreement Is Large at the Item Level, and Hidden by Small Leaderboards
#617156
arXiv (OAI Expanded)
Whose Gold? Annotator-Pool Disagreement Is Large at the Item Level, and Hidden by Small Leaderboards
#617656
arXiv (All)
Step-Level On-Policy Distillation: Interpolating Between On-Policy Distillation and Supervised Fine-Tuning
#619054
arXiv (OAI Expanded)
Step-Level On-Policy Distillation: Interpolating Between On-Policy Distillation and Supervised Fine-Tuning
#620733
arXiv (All)
Beyond BFI: The CSI for Enhanced Reliability and Validity in Evaluating LLM Personality Traits
#627905
arXiv (OAI Expanded)
Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades
#628149
arXiv (OAI Expanded)
Beyond BFI: The CSI for Enhanced Reliability and Validity in Evaluating LLM Personality Traits
#629767
arXiv (All)
Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades
#630011
arXiv (All)
Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering
#650870
arXiv (OAI Expanded)
Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering
#652323
arXiv (All)
Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck
#656807
arXiv (OAI Expanded)
Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification
#144200
arXiv (OAI)
Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning
#149278
arXiv (OAI)
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression
#163063
arXiv (OAI)
All Entities are Not Created Equal: Examining the Long Tail for Ultra-Fine Entity Typing
#173230
arXiv (OAI)
Can LLM Agents Simulate Multi-Turn Human Behavior? Evidence from Real Online Customer Behavior Data
#179138
arXiv (OAI)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
#179162
arXiv (OAI)
← Previous
Page 14 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.