arXiv:2608.26263v1 Announce Type: new Abstract: Large Language Models (LLMs) increasingly act as autonomous agents executing complex, long-running procedural skills.…
Existing agent runtimes maintain execution by continually appending ob…
Agent Mesh: Reliability Primitives for Non-Idempotent Agent Delegation - Identity Adequacy and Evidence Adequacy
arXiv:2608.26225v1 Announce Type: new Abstract: Autonomous agents increasingly perform bounded software tasks under an orchestrator that retries, resumes, and budgets them.…
The machinery such orchestrators reach for is the service mesh's: retr…
We report a failure study of a production agentic software-delivery pl…
arXiv:2608.26200v1 Announce Type: new Abstract: Modern video games combine first-person perception, rapid visual changes, persistent world state, and heterogeneous native controls.…
Existing game agents map visual and task context directly to actions b…
World-Action Models (WAMs) unify these objectives, but remain largely …
Agentic AI for operating scientific instruments for nanoscale characterization
arXiv:2608.26198v1 Announce Type: new Abstract: Operating a scientific instrument such as an atomic force microscope (AFM) requires continuous expert decision-making.…
A trained user defines the experimental intent, translates it into ins…
Existing automation usually addresses only parts of this workflow thro…
Is Your Neighborhood Safe? Place-based Stigma in Large Language Models' Urban Safety Judgments
arXiv:2608.26188v1 Announce Type: new Abstract: Large language models are increasingly used to inform safety decisions in cities, such as where it is safe to walk, rent, or travel.…
We ask whether such judgments track measured risk or the patterns atta…
We probe seven instruct-tuned models under three conditions that disso…
arXiv:2608.26178v1 Announce Type: new Abstract: There is growing interest in whether language models have stable preferences, for technical, safety, and philosophical reasons.…
We test 20 language models and find a range of preferences---stable di…
We run three forced-choice experiments on revealed rather than stated …
A Safety-Gated Multimodal AI Backend for Mental-Health Support: Hierarchical State Representation, Conservative Risk Fusion, and Controlled Generation in Anian
arXiv:2608.…
26162v1 Announce Type: new Abstract: Safety-critical mental-health sup…
This paper presents Anian, a safety-gated multimodal AI backend for pe…
Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration
arXiv:2608.26151v1 Announce Type: new Abstract: Subscriber attrition is a costly, persistent challenge for telecommunications providers, with monthly churn of roughly 1.…
9% in mature markets eroding billions in revenue annually.
Predictive models can flag at-risk customers accurately, yet they are …