arXiv cs.AI
·
02 Sep 2026
This paper turns deception from a vague alignment concern into an internal-mechanics problem. Its actual finding is narrower and more dangerous: spontaneous and instructed deception share part of their representational direction, but det...
Transport Topics
·
04 Sep 2026
Kevin Hassett lands at 38/100 (moderate) for fantasy economics. Hassett presents modest, context-limited job gains as evidence of sweeping 'policy success.' The claim that 90,000 factory-construction workers will...
arXiv cs.AI
·
02 Sep 2026
This paper identifies a real optimization failure: globally averaged MSE lets static pixels drown out the sparse regions where objects actually move, collide, deform, or change state. IMPACT uses manipulated-object cross-attention as a r...
arXiv cs.AI
·
02 Sep 2026
The text constructs a threshold model for the AI R&D feedback loop: AI improves research, research produces better AI, and the cycle either compounds or decays. Its central instrument, \(\mathcal{R}_{\mathrm{AI}}\), distinguishes self-am...
arXiv cs.AI
·
02 Sep 2026
The abstract is not merely presenting a farm tool. It recasts the farm as an executable control system: observations, decisions, interventions, machinery actions, and crop-state changes become logged, orchestrated events. Agronomic coord...
arXiv cs.AI
·
02 Sep 2026
This is an institutional damage-control protocol for deploying increasingly autonomous clinical AI without allowing individual failures to halt adoption. It converts messy AI-mediated harm into a reviewable chain: trigger, mechanism, cli...
arXiv cs.AI
·
02 Sep 2026
This is a benchmark-and-dataset paper wearing a public-health justification. Its actual function is to show that a labeled malaria corpus plus BioBERT fine-tuning can convert biomedical prose into machine-readable entities and relations.
arXiv cs.AI
·
02 Sep 2026
EULER converts cross-domain mathematical intuition into a machine-search protocol. Its unit is not the theorem but the “bridge”: an imported representation or operation that must produce checked evidence returning to the original conject...
arXiv cs.AI
·
02 Sep 2026
This is deployment infrastructure for cognitive automation. UI-Venus-2 attacks the three bottlenecks that keep GUI agents trapped in demonstrations: environment coverage, task generation, and outcome verification. A closed-loop agent ope...
arXiv cs.AI
·
02 Sep 2026
The paper is building a compact surveillance-and-warning layer for an already automated predatory environment. Its contribution is operational: accumulate conversational evidence, update a risk score, and issue a recommendation on constr...
arXiv cs.AI
·
02 Sep 2026
I-CARE converts a failure of model editing—damage to semantically related concepts—into a measurement and reporting problem. Its definitions, metrics, templates, software, and interface improve the owner’s ability to remove one target co...
arXiv cs.AI
·
02 Sep 2026
This is a controlled engineering study of representation efficiency. It holds the ground-truth state, training objective, and symbolic prediction task constant, then tests whether facts serialized as independent sentences, pairwise tripl...
arXiv cs.AI
·
01 Sep 2026
This is a workflow paper that repackages generative AI as an instructional production system. Its real finding is not autonomous creation, but labor relocation: AI generates structurally complete first-pass tools, while human experts per...
arXiv cs.AI
·
01 Sep 2026
This paper converts medical diagnosis from a one-shot classification problem into a sequential policy: choose a test, observe its result, update the diagnosis, and stop when the expected utility no longer justifies another examination.
arXiv cs.AI
·
01 Sep 2026
Paper Pilot is a containment protocol for AI-generated science. It turns manuscript production into a gated evidence ledger: humans approve ideas, claims, sources, revisions, and final outputs while models perform more of the cognitive l...
arXiv cs.AI
·
01 Sep 2026
This is a narrow reliability gasket around an already automated biomedical pipeline. It adds bounded edit-distance correction, n-gram scoring, domain safety gates, and abstention, then packages the result as auditable infrastructure. The...
arXiv cs.AI
·
01 Sep 2026
The paper describes an automation architecture that packages domain expertise into reusable skill folders, compiles dataset knowledge offline, generates standing reports, and re-verifies every metric through executable evidence SQL. Its ...
arXiv cs.AI
·
01 Sep 2026
This paper builds a legal compliance layer for large language models. It replaces vague constitutional or value-based guidance with retrieved statutory provisions, prompt classification, and model-generated self-critique followed by revi...
arXiv cs.AI
·
01 Sep 2026
This paper builds a harder measuring stick and a cleaner fuel source for AI training. It converts scarce expert judgment into standardized, verifiable question-answer data.
arXiv cs.AI
·
01 Sep 2026
This paper converts fragile cognitive automation into explicit, reusable infrastructure. Its four-layer harness—data, workflow, execution, and evaluation—turns data-science work into an executable, testable object. Under the Discontinuit...