Topic
i-2-7
Knowledge-graph topic · documents ABOUT i-2-7 across the archive
Documents about i-2-7
- Prior Audit-Repair Context Shifts LLM Verifier Thresholds Toward Leniency#617677arXiv (All)
- Architecture-Dependent Causal Transfer of Activation States Across Large Language Models#619068arXiv (OAI Expanded)
- Architecture-Dependent Causal Transfer of Activation States Across Large Language Models#620747arXiv (All)
- Consilience: Conformally Calibrated Communication Control for Hidden-Profile Multi-Agent Reasoning#656992arXiv (OAI Expanded)
- Consilience: Conformally Calibrated Communication Control for Hidden-Profile Multi-Agent Reasoning#659128arXiv (All)
- Evidence-Consistent Generative Detection under Scenario-Level Distribution Shift#661172arXiv (OAI Expanded)
- Evidence-Consistent Generative Detection under Scenario-Level Distribution Shift#662222arXiv (All)
- Evaluating Perspectival Biases in Cross-Modal Retrieval#789499arXiv (OAI Expanded)
- MedVision: Benchmarking Quantitative Medical Image Analysis#789516arXiv (OAI Expanded)
- Evaluating Perspectival Biases in Cross-Modal Retrieval#793503arXiv (All)
- MedVision: Benchmarking Quantitative Medical Image Analysis#793520arXiv (All)
- A Survey on Rubric-Guided Reinforcement Learning for Language Models#797691arXiv (OAI Expanded)
- A Survey on Rubric-Guided Reinforcement Learning for Language Models#799114arXiv (All)
- Decomposing Wrong-Consensus Agreement in LLM Self-Consistency#800142arXiv (All)
- MABPD: Multi-Agent Bias Probing & Detection via Structured Argument Debate#971506arXiv (All)
- Policy-Grounded Safety Evaluation of 20 Large Language Models#183283arXiv (OAI)
- ICE: Intervention-Consistent Explanation Evaluation with Statistical Grounding for LLMs#919831arXiv (OAI Expanded)
- ICE: Intervention-Consistent Explanation Evaluation with Statistical Grounding for LLMs#924060arXiv (All)
- CytoNet: A Foundation Model for the Human Cerebral Cortex at Cellular Resolution#668353arXiv (OAI Expanded)
- CytoNet: A Foundation Model for the Human Cerebral Cortex at Cellular Resolution#670418arXiv (All)
- 7 U.S.C. § 1426 — Repealed. Pub. L. 104–127, title I, § 171(b)(2)(I), Apr. 4, 1996, 110 Stat. 938#561774US Code (LII)
- 7 U.S.C. § 1433f — Repealed. Pub. L. 104–127, title I, § 171(b)(2)(I), Apr. 4, 1996, 110 Stat. 938#561796US Code (LII)
- Measuring AI Accountability Through Argumentation Analysis: Can Model Reasoning Withstand Scrutiny?#972039arXiv (All)
- Share the Judge, Learn the Deferral: Where Specialization Helps LLM Evaluation#650962arXiv (OAI Expanded)
- Share the Judge, Learn the Deferral: Where Specialization Helps LLM Evaluation#652415arXiv (All)
- UpgradeBench: A Decision-Centric Benchmark for Upgrading Fine-Tuned LLM Specialists#661060arXiv (OAI Expanded)
- No Judgment Without a Reason: Counterfactual Receipts for Versioned AI Evaluators#661078arXiv (OAI Expanded)
- UpgradeBench: A Decision-Centric Benchmark for Upgrading Fine-Tuned LLM Specialists#662110arXiv (All)
- No Judgment Without a Reason: Counterfactual Receipts for Versioned AI Evaluators#662128arXiv (All)
- TW-LegalBench: Measuring Taiwanese Legal Understanding#668576arXiv (OAI Expanded)
- TW-LegalBench: Measuring Taiwanese Legal Understanding#670641arXiv (All)
- Compile, Don't Memorize: A Context Compilation Architecture (CCA) for In-Context Learning#820771arXiv (All)
- Are Verifier Errors Independent Within a GRPO Group? Evidence from Qwen2.5 Rollouts#973967arXiv (All)
- SAGE: A Hierarchical Framework for Evaluating Interpretive Literary Quality in Narratives#974718arXiv (All)
- AgentIdeaBench: Benchmarking Scientific Ideation in the Agent Era#984521arXiv (All)
- Proof-Carrying Cognition: Closing the Verification Gap with Reality-Settled Reward#996598arXiv (All)
- LLM-as-a-Judge Is Not an Oracle: Why Self-Improving Agents Need Deterministic Guardrails#822636arXiv (All)
- Personality Without Persons? A Psychometric Critique of Big Five Testing in Large Language Models#919879arXiv (OAI Expanded)
- Personality Without Persons? A Psychometric Critique of Big Five Testing in Large Language Models#924108arXiv (All)
- Local derivations is a Lie algebra#986971arXiv (All)
- MELD: A Protocol for Merging Knowledge Across Distributed Agentic Memories#619078arXiv (OAI Expanded)
- MELD: A Protocol for Merging Knowledge Across Distributed Agentic Memories#620757arXiv (All)
- ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation#797655arXiv (OAI Expanded)
- ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation#799078arXiv (All)
- 7 U.S.C. § 1441-2 — Repealed. Pub. L. 104–127, title I, § 171(b)(2)(A), Apr. 4, 1996, 110 Stat. 938#561730US Code (LII)
- 7 U.S.C. § 1444-2 — Repealed. Pub. L. 104–127, title I, § 171(b)(2)(B), Apr. 4, 1996, 110 Stat. 938#561736US Code (LII)
- RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation#664381arXiv (OAI Expanded)
- RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation#665881arXiv (All)
- BiasMix-Finance: Post-Generation KYC Guardrails for LLM Portfolio Advice#788892arXiv (OAI Expanded)
- BiasMix-Finance: Post-Generation KYC Guardrails for LLM Portfolio Advice#792896arXiv (All)
Topic record · derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.