微信内可能无法直接打开本站。请点右上角 ··· → 在浏览器打开 ,或复制链接后用系统浏览器访问。
综合
官方
企业
汇聚
1 / 24
Beyond the Best Guess: Improving LLM Solution Coverage with Evolution Strategies
arXiv:2608.12679v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in discovery domains such as math and science.…
The usual approach is to present the problem to the model and use its …
However, beyond this best guess, discovery can be enhanced by increasi…
In a process called pass@k, the model is allowed to explore the soluti…
RSS 官方收录 · 可信分层展示
The Role of Natural Language Understanding in Multimodal Video-Based Dengue Diagnosis
arXiv:2608.…
12677v1 Announce Type: new Abstract: Detecting infection-related behav…
In this study, a YOLO- and Contrastive Language-Image Pre-training (CL…
First, YOLO is used to isolate mosquito regions from the background.
RSS 官方收录 · 可信分层展示
Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs
arXiv:2608.…
12675v1 Announce Type: new Abstract: Retrieval-Augmented Generation (R…
Existing privacy research on RAG has focused on preventing unauthorize…
However, another important problem that is often overlooked in RAG pri…
RSS 官方收录 · 可信分层展示
Lines and Ladders: A Context-Aware Multi-Agent Framework for Large-Scale Retail Price Taxonomy
arXiv:2608.12674v1 Announce Type: new Abstract: Maintaining price consistency and executing an Every Day Low Price strategy is critical for global retailers.…
However, with catalogs spanning millions of active items, manual gover…
Inconsistent pricing across item variants distorts customer value perc…
To address this, we present a scalable, context-aware Multi-Agent Fram…
RSS 官方收录 · 可信分层展示
On the Expressive Power of Transformers
arXiv:2608.12671v1 Announce Type: new Abstract: Multi-layer transformers form the critical component of essentially all large language models (LLMs) in use today.…
Because of their ubiquity and computational capability, there is a rap…
In this endeavor, circuit complexity has by and large emerged as the "…
Here, we present an overview of selected results that delineate the ex…
RSS 官方收录 · 可信分层展示
Noah Reimagines the Converse Jack Purcell 1935 in Plaid
Name: Noah x Converse Jack Purcell 1935Colorway: TBCSKU: TBCMSRP: ¥18,700 JPY (approx.…
$117 USD)Release Date: August 21Where to Buy: Noah, ConverseNoah has t…
In celebrating the Jack Purcell 1935’s 90th anniversary, the sneaker i…
Noah adds its own identity through an original plaid pattern on the up…
RSS 官方收录 · 可信分层展示
鹿明发布MOS2:全球首个双臂负载50kg轮臂式机器人,加速AI Worker进入产业现场
8月14日,鹿明机器人发布全新重载轮臂式具身智能机器人Lumos MOS2。面向真实工业场景,Lumos MOS2具备50kg双臂负载能力,同时在硬件性能、全向移动能力、多模态感知系统和整机控制架构等方面实现全面升级,胜任真实工业场景高强度、持续性作业任务。…
鹿明机器人创始人兼CEO喻超表示,具身智能正在加速进入产业落地阶段,如何让机器人更高效地完成更多真实任务,取决于数据获取效率、硬件成本及产业…
在这一判断下,MOS2被定位为面向工业场景的重载AI Worker:既拥有7×24小时不间断工作的强健躯体,也具备环境感知、任务理解与自主操…
当前工业现场中,仍存在大量同时需要负载能力、灵活移动和复杂操作的任务,而这一类场景仍缺少成熟的具身智能机器人解决方案。
RSS 官方收录 · 可信分层展示
索塔无界:全球首家原生物理世界模型落地商超,具身智能迎来“索塔时刻”
在具身智能仍陷于“资本热、落地冷”之际,成立仅4个月的索塔无界,率先以欧洲最大商超集团的战略合作,打破商业落地僵局。根据规划,双方未来3年将在真实商超场景中部署超过千台具身智能机器人。…
索塔无界不造硬件,而是为机器人打造“原生物理大脑”——首次将4D世界动作模型部署于真实商超货架之间,开启物理智能的“场景收敛+通用泛化”新路径。
破局:不追“虚拟温床”,直击“物理考场”2026年的具身智能,融资额屡创新高,但机器人进工厂仍困于专机专用、进家庭做家务仍在可望难及的远期蓝图。
多数玩家沉迷视频生成或仿真训练,这些“虚拟温床”对具身操作并不实质。
RSS 官方收录 · 可信分层展示
发布即热销:大疆 Osmo 360 II 首日拿下全渠道销量 TOP1
8月13日20:00,大疆正式发布全新 8K 全景相机 Osmo 360 II。根据24小时战报数据,Osmo 360 II 全渠道首日销量位列 TOP1,并在京东、天猫、抖音三大平台全景相机相关榜单中登顶,上市首日即展现出强劲市场表现。…
战报显示,Osmo 360 II 发布4小时后,已拿下京东、天猫、抖音三大平台新品相关榜单 TOP1;截至首发24小时,产品热度进一步释放,…
与此同时,Osmo 360 II 上市首日全网总曝光量达到 1.
17 亿+,销售端与传播端同时展现出强劲首发势能。
RSS 官方收录 · 可信分层展示
Frozen 3 Trailer Reveals Anna’s Royal Wedding and A New Icy Magic
Frozen 3 Trailer Reveals Anna’s Royal Wedding and A New Icy Magic
Frozen 3 Trailer Reveals Anna’s Royal Wedding and A New Icy Magic
RSS 官方收录 · 可信分层展示
Designing AI Pipelines for Decision-Ready ITSM Intelligence
arXiv:2608.…
12670v1 Announce Type: new Abstract: IT service management (ITSM) syst…
This paper presents a sociotechnical AI pipeline, designed and evaluat…
The pipeline combines LLM-based schema normalization, HDBSCAN sub-topi…
RSS 官方收录 · 可信分层展示
General Probabilities of Causation with Causal Knowledge
arXiv:2608.…
12657v1 Announce Type: new Abstract: Probabilities of causation (PoCs)…
Tian and Pearl first derived theoretically sharp bounds for binary PoC…
Mueller et al.
RSS 官方收录 · 可信分层展示
SteerBench-Work: A Benchmark for Agent Steering at Action Boundaries
arXiv:2608.12654v1 Announce Type: new Abstract: Long-running LLM agents act through tools, and a single step can send an email, merge a pull request, or wire a payment.…
The steering decision is the pre-commit choice at that boundary: proce…
We introduce SteerBench-Work, an incident-anchored, bidirectional benc…
Release v2026-05 contains 106 scenarios anchored in public incidents, …
RSS 官方收录 · 可信分层展示
@skills: Attention is all you have
arXiv:2608.12610v1 Announce Type: new Abstract: There are 56,804 public agent skills today, and teams write many more privately.…
The dominant delivery model is installation: once installed, a skill's…
This leaves the long tail with no practical path to use and forces tea…
We observe that installation bundles three separable functions: conten…
RSS 官方收录 · 可信分层展示
Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence
arXiv:2608.12645v1 Announce Type: new Abstract: LLM judges have become central infrastructure for model evaluations, online grading, and reward modeling.…
Judges are typically validated by accuracy on golden data, but accurac…
We introduce the \emph{Wiggle Framework}, a unified stress test for ep…
The framework decomposes judge robustness along three dimensions: Mech…
RSS 官方收录 · 可信分层展示
Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues
arXiv:2608.…
12599v1 Announce Type: new Abstract: Multi-turn dialogues let users re…
No existing instrument measures this influence per clause, predicts it…
\sysname{} closes the three gaps through the model API alone: a contra…
RSS 官方收录 · 可信分层展示
DiG-bench: Discovery in Games
arXiv:2608.12593v1 Announce Type: new Abstract: Discovery---formulating novel generalizations---is a central part of the scientific process.…
Despite its importance, there is a gap in the current AI benchmark lan…
To address this gap, we release a new benchmark: DiG-bench (Discovery …
DiG-bench consists of a set of 70 independent games.
RSS 官方收录 · 可信分层展示
Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting
arXiv:2608.…
12590v1 Announce Type: new Abstract: Thyroid ultrasound diagnosis requ…
We present ThyroidXAgent, a clinician-interactive agentic AI system th…
The system was developed using OpenThyroidDB, a multicentre, multitask…
RSS 官方收录 · 可信分层展示
Reasoning Jury: Multi-Model Consensus for Evaluating Reasoning Traces
arXiv:2608.…
12585v1 Announce Type: new Abstract: Improving reasoning LLMs requires…
Additionally, surfacing reasoning mistakes that the model makes would …
Due to the difficulty of this complex task on long reasoning traces, s…
RSS 官方收录 · 可信分层展示
Trie Automata for Constrained Decoding over Large Finite Sets
arXiv:2608.…
12574v1 Announce Type: new Abstract: Large language models increasingl…
Current constrained decoding systems handle this through general-purpo…
We introduce the trie automaton, a specialized mechanism that exploits…
RSS 官方收录 · 可信分层展示
CAS: A Causal Attribution Score for Local and Global Explainable Artificial Intelligence
arXiv:2608.…
12555v1 Announce Type: new Abstract: Predictive explanation methods at…
We introduce the Causal Attribution Score (CAS), a compact score archi…
CAS starts from an identified interventional coalition game, allocates…
RSS 官方收录 · 可信分层展示
$\varepsilon$-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution
arXiv:2608.…
12522v1 Announce Type: new Abstract: LLM-based program evolution syste…
We introduce $\varepsilon$-MemEvo, a framework for cross-task knowledg…
$\varepsilon$-MemEvo stores prior experience as task-agnostic tactic m…
RSS 官方收录 · 可信分层展示
Governed Persistent Memory: Source-Bound State Semantics and Fail-Closed Release for Long-Horizon Agents
arXiv:2608.…
12476v1 Announce Type: new Abstract: Long-term agent memory is usually…
We introduce Governed Persistent Memory (GPM), an auditable bitemporal…
Five executable clauses cover ledger integrity, source binding, confli…
RSS 官方收录 · 可信分层展示
MindMemOS: A Portable and Self-Evolving Memory Operating Layer for AI Agents
arXiv:2608.…
12428v1 Announce Type: new Abstract: Memory is a core component of AI …
However, existing memory systems often remain fixed after development,…
We present MindMemOS, a portable and self-evolving memory operating la…
RSS 官方收录 · 可信分层展示
上滑下一条
上滑 · j/k · m/u · h 隐藏 · a 稍后 · o 原文 · e 详情 · t 今日 · i 模式 · f 搜索