微信内可能无法直接打开本站。请点右上角 ··· → 在浏览器打开 ,或复制链接后用系统浏览器访问。
综合
官方
企业
汇聚
当前信源:arXiv cs.AI
· 清除信源筛选
Bioinfoysis Technical Report
arXiv:2609.…
03871v1 Announce Type: new Abstract: Large language model agents have …
This design is poorly suited to long-horizon bioinformatics tasks, whe…
RSS 官方收录 · 可信分层展示
STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation
arXiv:2609.…
03874v1 Announce Type: new Abstract: Retrieval Augmented Generation (R…
LLMs are improving at handling long context, but still suffer from "lo…
RSS 官方收录 · 可信分层展示
Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations
arXiv:2609.03860v1 Announce Type: new Abstract: Retail supply chain operations rely on coupled decision modules that must adapt as requirements evolve.…
LLMs offer a natural-language interface for this task, but existing me…
Extending them to heterogeneous decision pipelines is challenging beca…
RSS 官方收录 · 可信分层展示
Semantic Bayesian World Models
arXiv:2609.…
03834v1 Announce Type: new Abstract: Knowledge graphs describe reality…
We argue that this mismatch is why the integration of language models …
RSS 官方收录 · 可信分层展示
CauseCollab: Causal Unified and Modality-Agnostic Network for Heterogeneous Collaborative Perception
arXiv:2609.…
03818v1 Announce Type: new Abstract: Collaborative perception enhances…
Recent protocol-based two-stage methods alleviate this problem by mapp…
RSS 官方收录 · 可信分层展示
SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation
arXiv:2609.…
03806v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) ge…
Progress, however, is held back by the lack of domain-specific evaluat…
RSS 官方收录 · 可信分层展示
Govern the Model, Not Only the Data: Storage, Circulation, and Learning in Creative AI
arXiv:2609.…
03800v1 Announce Type: new Abstract: Federated learning is increasingl…
It borrows the vocabulary of the federated social web, yet inverts its…
RSS 官方收录 · 可信分层展示
Transfiver: Human-AI Co-Inference through a Shared Editable State
arXiv:2609.…
03797v1 Announce Type: new Abstract: Long-term human-AI interaction is…
We introduce the TRANSparent Framework for Interactive, Verifiable, Ed…
RSS 官方收录 · 可信分层展示
DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions
arXiv:2609.…
03787v1 Announce Type: new Abstract: AI agents increasingly gather evi…
A final output alone cannot show which evidence, tool state, rule, aut…
RSS 官方收录 · 可信分层展示
Rethinking World Models for Safety-Critical Embodied Systems
arXiv:2609.…
03774v1 Announce Type: new Abstract: World models have progressed from…
However, high predictive likelihood and visual fidelity do not necessa…
RSS 官方收录 · 可信分层展示
SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation
arXiv:2609.…
03753v1 Announce Type: new Abstract: As large language models (LLMs) b…
We introduce SimSkill, a self-evolving agent built around the Simulati…
RSS 官方收录 · 可信分层展示
Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation
arXiv:2609.…
03727v1 Announce Type: new Abstract: Large language model agents can p…
Proactive service moves the decision upstream: an agent must infer ser…
RSS 官方收录 · 可信分层展示
Artificial Intelligence for Energy Optimization in Data Centers
arXiv:2609.03716v1 Announce Type: new Abstract: Data centers are increasingly optimized by artificial intelligence and, at the same time, increasingly loaded by it.…
The literature treats these as two unrelated problems: control studies…
We screen roughly 194 papers retrieved through a documented protocol, …
RSS 官方收录 · 可信分层展示
Counterfactual Routing Using Integer Programming with Constraint Generation
arXiv:2609.03707v1 Announce Type: new Abstract: We present our submission to the IJCAI 2025 'Counterfactual Routing Competition' (CRC 25).…
The goal of the competition is to find counterfactual explanations for…
This requires deciding what the minimal changes to a road network woul…
RSS 官方收录 · 可信分层展示
Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study
arXiv:2609.03702v1 Announce Type: new Abstract: General-purpose code embeddings power tools for code search, classification, and retrieval.…
Compact transformer encoders for code typically rely on either human-w…
We empirically study an alternative: contrastive pretraining of small …
RSS 官方收录 · 可信分层展示
Analysis of Prompt Engineering for Drug Toxicity Prediction
arXiv:2609.03635v1 Announce Type: new Abstract: Clinical trials in the UK can cost up to {\pounds}1.…
3 million, with approximately 90% drug failure rate.
Toxicity is a major contributing factor in drug failure.
RSS 官方收录 · 可信分层展示
A computable representation of the physical laboratory enables verifiable workflows
arXiv:2609.…
03621v1 Announce Type: new Abstract: Making science computable require…
A computable representation of the physical laboratory is established …
RSS 官方收录 · 可信分层展示
KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents
arXiv:2609.…
03588v1 Announce Type: new Abstract: As LLMs increasingly act through …
We introduce KC-Bench, a controlled multi-turn benchmark for measuring…
RSS 官方收录 · 可信分层展示
The Attention Triangle in Audio-Video Models
arXiv:2609.…
03586v1 Announce Type: new Abstract: Audio-video diffusion models rely…
We study these models by probing and analyzing the ``attention triangl…
RSS 官方收录 · 可信分层展示
HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews
arXiv:2609.…
03580v1 Announce Type: new Abstract: The growing scale of academic pee…
Existing hallucination benchmarks are not designed for peer review, wh…
RSS 官方收录 · 可信分层展示
GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis
arXiv:2609.…
03553v1 Announce Type: new Abstract: Policy analysis requires more tha…
LLM-based policy simulations model these processes at scale, but their…
RSS 官方收录 · 可信分层展示
Dalek: A Constructive Agent Machine
arXiv:2609.…
03546v1 Announce Type: new Abstract: We present Dalek, a closed machin…
The machine is built from three primitives---actors, messages, and cha…
RSS 官方收录 · 可信分层展示
Feature Reconfiguration With Visual Prior for Medical Lesion Segmentation
arXiv:2609.03535v1 Announce Type: new Abstract: Lesion segmentation in medical images plays a critical role in clinical diagnosis and treatment planning.…
Despite significant advances, lesion segmentation remains challenging …
Existing encoder-decoder based methods mainly focus on enhancing featu…
RSS 官方收录 · 可信分层展示
NeoRed: A Knowledge-Logic-Alignment Multimodal Large Language Model for Neonatal Respiratory Disease Diagnosis
arXiv:2609.…
03527v1 Announce Type: new Abstract: Neonatal respiratory diseases are…
Despite recent advances, existing Multimodal Large Language Models (ML…
RSS 官方收录 · 可信分层展示
CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning
arXiv:2609.…
03526v1 Announce Type: new Abstract: Multimodal language models achiev…
To probe this distinction, we introduce CulturalMenuBench, a benchmark…
RSS 官方收录 · 可信分层展示
What Matters for Aggressive Decoding-Time KV Eviction? Temporal Aggregation and Ranking Preservation
arXiv:2609.…
03515v1 Announce Type: new Abstract: Decoding-time KV cache compressio…
Under aggressive KV compression, we find that exponential-moving-avera…
RSS 官方收录 · 可信分层展示
PPO-STGNN: A Proximal Policy Optimization Approach with Spatio-Temporal Graph Neural Networks for DAG Task Scheduling in Cloud-Edge-End Computing
arXiv:2609.…
03503v1 Announce Type: new Abstract: With the rapid development of the…
However, cloud, edge, and end nodes are highly heterogeneous in comput…
RSS 官方收录 · 可信分层展示
GrowPage: On-Demand KV Budgeting for Efficient LLM Reasoning Serving
arXiv:2609.03494v1 Announce Type: new Abstract: Long-output reasoning has made the key--value (KV) cache a critical memory bottleneck for efficient LLM serving.…
Existing KV compression methods usually rely on a predefined per-reque…
However, reasoning workloads exhibit substantial demand variation: dif…
RSS 官方收录 · 可信分层展示
AutoGraphForge: Towards Automated Graph Theory Discovery
arXiv:2609.…
03478v1 Announce Type: new Abstract: We report on our ongoing project …
Conjecture generation is counterexample-guided and runs in rounds: a G…
RSS 官方收录 · 可信分层展示
Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Language Models
arXiv:2609.…
03493v1 Announce Type: new Abstract: Modern vision-language models (VL…
To acquire this missing evidence, agentic VLMs invoke tools such as im…
RSS 官方收录 · 可信分层展示
下滑 · j/k · m/u · i 信流 · a 稍后 · o 原文 · e 详情 · t 今日 · f 搜索