Skip to main content

XMT

短闻

信流 · 上滑连读 · 来源可核

今日 稍后 搜索 RSS

当前信源:arXiv cs.AI · 清除信源筛选

1 / 24
Aggregate arXiv cs.AI 人工智能 45″

Conformity Mitigations in Large Language Models Lie on a Single Resistance-Receptivity Frontier

arXiv:2608.…

  • 11247v1 Announce Type: new Abstract: Recent advances in language model…
  • Each agent sees what the others assert before it answers, so peer opin…
  • We measure that displacement in 23 open-weight models, 19 conditions, …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Towards the Harness of Embodied Agents

arXiv:2608.…

  • 11246v1 Announce Type: new Abstract: The success of coding agents has …
  • We ask whether the same paradigm extends to embodied agents in the phy…
  • We present Thea, a harness in which an agentic loop orchestrates robot…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

BEST-KAG: Enhancing Question Answering of Building Engineering Standards with Multimodal Knowledge Graph Modeling and Large Language Model

arXiv:2608.11244v1 Announce Type: new Abstract: Construction standards are critical for building safety and sustainability.…

  • Existing standard application workflows rely on keyword-based document…
  • To address these limitations, this study develops a multimodal knowled…
  • The framework introduces 1) a multimodal knowledge graph (MKG) for uni…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Towards Sustainable Learning in Online Education: A Reinforcement Learning Approach

arXiv:2608.…

  • 11245v1 Announce Type: new Abstract: Online education offers unprecede…
  • To address these challenges, we introduce AI Tutor, a reinforcement le…
  • In the short term, AI-Tutor draws on cognitive theory to guide learner…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

The Off-Support Barrier: Why Semantic Safety Constraints Are Not Learning-Problem Invariants, and What Follows for Prior Design, Containment, and Verification

arXiv:2608.…

  • 11243v1 Announce Type: new Abstract: We argue that a single structural…
  • , the agent does not escape its sandbox) is an off-support object.
  • Formally, if q is the data distribution and \(p(\cdot\mid w)\) the mod…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

RecSys Factory: Bounding LLM Agent Autonomy to Decision Points in the Industrial Recommender Lifecycle

arXiv:2608.…

  • 11241v1 Announce Type: new Abstract: Deploying LLM agents into industr…
  • Any two can be maximized against the third.
  • We present RecSys Factory, an LLM-agent platform deployed for 78 days …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

VQ-bench: A Composable Vector Quantization Framework

arXiv:2608.11240v1 Announce Type: new Abstract: Vector quantization is an old problem but has recently become central to AI infrastructure.…

  • It is therefore experiencing a surge of renewed engineering and resear…
  • This paper provides a unified framework for developing and benchmarkin…
  • We describe 7 common conceptual quantization primitives and show how t…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability

arXiv:2608.…

  • 11238v1 Announce Type: new Abstract: Retrieval-augmented generation im…
  • We propose Q-CARE, a query-agnostic and fully reference-free framework…
  • Q-CARE establishes a unified evaluation principle based on query cover…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

CORA-Diff: Confidence-Oriented Residual Acceptance for Efficient Diffusion Language Model Inference

arXiv:2608.11235v1 Announce Type: new Abstract: Diffusion language models (DLMs) update many tokens in parallel, yet practical decoders often use a fixed denoising horizon.…

  • Many predictions stabilize early, but blockwise decoding continues unt…
  • Existing accelerators often rely on learned filters, modified scores, …
  • We ask whether native trajectory signals can identify residual positio…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Geometry-aware Incremental Neural Operator for Long-Horizon PDE prediction

arXiv:2608.11237v1 Announce Type: new Abstract: Neural operators have shown strong potential for learning solution operators of partial differential equations (PDEs).…

  • However, long-horizon autoregressive prediction remains challenging: l…
  • Existing methods mainly improve state representations and operator bac…
  • To address these issues, we propose a geometry-aware incremental neura…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk

arXiv:2608.11234v1 Announce Type: new Abstract: Managing modern computing infrastructure has become a steadily harder problem due to the ever-increasing complexity.…

  • Recent advances in AI agents create a timely opportunity to automate i…
  • We present InfraBench, a benchmark suite for evaluating AI agents on r…
  • Experiments with 15 agent-model configurations show that even the stro…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

LinearKV: One Cached State Suffices for Position-Independent Caching in Hybrid LLMs

arXiv:2608.11231v1 Announce Type: new Abstract: LLM serving is increasingly accelerated by position-independent caching (PIC).…

  • Existing PIC methods, however, are built for full-attention models, wh…
  • Hybrid LLMs break these primitives---they replace most attention layer…
  • This raises a natural question: can PIC benefit hybrid models, and wha…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

The Edge-based Contiguous p-median Problem with Connections to Logistics Districting

arXiv:2608.…

  • 11230v1 Announce Type: new Abstract: This paper introduces the edge-ba…
  • Two binary programming models are introduced, both of which incorporat…
  • The first model requires an exponential number of cut set-based constr…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Forecasting Side Effects of Activation Steering

arXiv:2608.…

  • 11227v1 Announce Type: new Abstract: Activation steering modifies a la…
  • While effective, steering often produces unintended side effects on ot…
  • We therefore ask: can these side effects be forecasted before steering…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended Version)

arXiv:2608.…

  • 11229v1 Announce Type: new Abstract: Comparative feedback, asking peop…
  • Preference-based reward learning typically casts the human teacher as …
  • We argue this forfeits the teacher's defining advantage: knowledge of …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Cutting AI Datacenter Energy with Reinforcement Learning: Measured Power Control of LLM Training from One GPU to the Fleet

arXiv:2608.…

  • 11226v1 Announce Type: new Abstract: Reinforcement-learning post-train…
  • We instrument GRPO training with half-second power telemetry at 7B, 14…
  • Against the full 500-step 7B trace, the controller cuts power-limit vi…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Harnessing agent memory to build lifelong AI partners for materials scientists

arXiv:2608.…

  • 11224v1 Announce Type: new Abstract: Materials research advances throu…
  • This experience is essential for reproducibility and knowledge transfe…
  • Here we argue that a lifelong AI partner for materials science can be …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Identity from the Outside: A Conceptual Framework and Research Program for AI Personality Clones

arXiv:2608.11225v1 Announce Type: new Abstract: AI "personality clones" force a re-examination of personal identity in operational terms.…

  • Setting aside the hard problem of consciousness, we approach identity …
  • We distinguish three criteria that "identity" conflates: fidelity to a…
  • We propose a six-term factorization of observed identity (substrate, d…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

A Conceptual Framework for Refining Influence Knowledge from Simulation Evidence in Cyber-Physical Systems

arXiv:2608.…

  • 11221v1 Announce Type: new Abstract: Cyber-physical systems (CPS) are …
  • The behaviour of these systems emerges from the interaction between th…
  • Simulation and co-simulation have become essential approaches for anal…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

LLMs in Process Diagram Engineering: From Optimal PFDs to Validated P&IDs

arXiv:2608.…

  • 11220v1 Announce Type: new Abstract: Nowadays, the creation of a proce…
  • Applying artificial intelligence in the task could potentially lead no…
  • This research presents P&ID Pilot - a practical end-to-end AI pipeline…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

From Monolithic to Modular: Segment-level Automatic Prompt Optimization

arXiv:2608.11219v1 Announce Type: new Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others.…

  • We present SAPO, a segment-level APO method that decomposes prompts in…
  • The optimization loop uses one LLM with static meta-prompts and struct…
  • We describe a train/validation protocol and a two-stage generation pro…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

MaSRead: Content-Addressed Reading of Replicated Latent Stores

arXiv:2608.11218v1 Announce Type: new Abstract: Independent agents that reason in latent space can share computed state as key-value cache fragments rather than text.…

  • Merged by a conflict-free replicated data type, these fragments form a…
  • Yet a later query, unknown at encode time, cannot reliably read the me…
  • MaSRead addresses the read to content.

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research

arXiv:2608.…

  • 11216v1 Announce Type: new Abstract: World modeling is an unsettled fi…
  • This makes it an ideal testbed for AI coding agents acting as autonomo…
  • We introduce AutoWorldModel-Bench, a closed-loop benchmark in which fr…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop

arXiv:2608.…

  • 11215v1 Announce Type: new Abstract: Simulating societies of many larg…
  • We turn a statistical-physics observation into a method: replace each …
  • Whether this works is decided before the simulation runs, chiefly by w…

RSS 官方收录 · 可信分层展示

详情 原文 分享图