Skip to main content

XMT

短闻

信流 · 上滑连读 · 来源可核

今日 稍后 搜索 RSS
1 / 24
Aggregate arXiv cs.AI 人工智能 45″

Beyond the Best Guess: Improving LLM Solution Coverage with Evolution Strategies

arXiv:2608.12679v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in discovery domains such as math and science.…

  • The usual approach is to present the problem to the model and use its …
  • However, beyond this best guess, discovery can be enhanced by increasi…
  • In a process called pass@k, the model is allowed to explore the soluti…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

The Role of Natural Language Understanding in Multimodal Video-Based Dengue Diagnosis

arXiv:2608.…

  • 12677v1 Announce Type: new Abstract: Detecting infection-related behav…
  • In this study, a YOLO- and Contrastive Language-Image Pre-training (CL…
  • First, YOLO is used to isolate mosquito regions from the background.

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs

arXiv:2608.…

  • 12675v1 Announce Type: new Abstract: Retrieval-Augmented Generation (R…
  • Existing privacy research on RAG has focused on preventing unauthorize…
  • However, another important problem that is often overlooked in RAG pri…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Lines and Ladders: A Context-Aware Multi-Agent Framework for Large-Scale Retail Price Taxonomy

arXiv:2608.12674v1 Announce Type: new Abstract: Maintaining price consistency and executing an Every Day Low Price strategy is critical for global retailers.…

  • However, with catalogs spanning millions of active items, manual gover…
  • Inconsistent pricing across item variants distorts customer value perc…
  • To address this, we present a scalable, context-aware Multi-Agent Fram…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

On the Expressive Power of Transformers

arXiv:2608.12671v1 Announce Type: new Abstract: Multi-layer transformers form the critical component of essentially all large language models (LLMs) in use today.…

  • Because of their ubiquity and computational capability, there is a rap…
  • In this endeavor, circuit complexity has by and large emerged as the "…
  • Here, we present an overview of selected results that delineate the ex…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate Hypebeast 时尚 45″

Noah Reimagines the Converse Jack Purcell 1935 in Plaid

Name: Noah x Converse Jack Purcell 1935Colorway: TBCSKU: TBCMSRP: ¥18,700 JPY (approx.…

  • $117 USD)Release Date: August 21Where to Buy: Noah, ConverseNoah has t…
  • In celebrating the Jack Purcell 1935’s 90th anniversary, the sneaker i…
  • Noah adds its own identity through an original plaid pattern on the up…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate 雷锋网 综合科技 45″

鹿明发布MOS2:全球首个双臂负载50kg轮臂式机器人,加速AI Worker进入产业现场

8月14日,鹿明机器人发布全新重载轮臂式具身智能机器人Lumos MOS2。面向真实工业场景,Lumos MOS2具备50kg双臂负载能力,同时在硬件性能、全向移动能力、多模态感知系统和整机控制架构等方面实现全面升级,胜任真实工业场景高强度、持续性作业任务。…

  • 鹿明机器人创始人兼CEO喻超表示,具身智能正在加速进入产业落地阶段,如何让机器人更高效地完成更多真实任务,取决于数据获取效率、硬件成本及产业…
  • 在这一判断下,MOS2被定位为面向工业场景的重载AI Worker:既拥有7×24小时不间断工作的强健躯体,也具备环境感知、任务理解与自主操…
  • 当前工业现场中,仍存在大量同时需要负载能力、灵活移动和复杂操作的任务,而这一类场景仍缺少成熟的具身智能机器人解决方案。

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate 雷锋网 综合科技 45″

索塔无界:全球首家原生物理世界模型落地商超,具身智能迎来“索塔时刻”

在具身智能仍陷于“资本热、落地冷”之际,成立仅4个月的索塔无界,率先以欧洲最大商超集团的战略合作,打破商业落地僵局。根据规划,双方未来3年将在真实商超场景中部署超过千台具身智能机器人。…

  • 索塔无界不造硬件,而是为机器人打造“原生物理大脑”——首次将4D世界动作模型部署于真实商超货架之间,开启物理智能的“场景收敛+通用泛化”新路径。
  • 破局:不追“虚拟温床”,直击“物理考场”2026年的具身智能,融资额屡创新高,但机器人进工厂仍困于专机专用、进家庭做家务仍在可望难及的远期蓝图。
  • 多数玩家沉迷视频生成或仿真训练,这些“虚拟温床”对具身操作并不实质。

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate 雷锋网 综合科技 45″

发布即热销:大疆 Osmo 360 II 首日拿下全渠道销量 TOP1

8月13日20:00,大疆正式发布全新 8K 全景相机 Osmo 360 II。根据24小时战报数据,Osmo 360 II 全渠道首日销量位列 TOP1,并在京东、天猫、抖音三大平台全景相机相关榜单中登顶,上市首日即展现出强劲市场表现。…

  • 战报显示,Osmo 360 II 发布4小时后,已拿下京东、天猫、抖音三大平台新品相关榜单 TOP1;截至首发24小时,产品热度进一步释放,…
  • 与此同时,Osmo 360 II 上市首日全网总曝光量达到 1.
  • 17 亿+,销售端与传播端同时展现出强劲首发势能。

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate CNET News 综合科技 6″

Frozen 3 Trailer Reveals Anna’s Royal Wedding and A New Icy Magic

Frozen 3 Trailer Reveals Anna’s Royal Wedding and A New Icy Magic

  • Frozen 3 Trailer Reveals Anna’s Royal Wedding and A New Icy Magic

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Designing AI Pipelines for Decision-Ready ITSM Intelligence

arXiv:2608.…

  • 12670v1 Announce Type: new Abstract: IT service management (ITSM) syst…
  • This paper presents a sociotechnical AI pipeline, designed and evaluat…
  • The pipeline combines LLM-based schema normalization, HDBSCAN sub-topi…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

General Probabilities of Causation with Causal Knowledge

arXiv:2608.…

  • 12657v1 Announce Type: new Abstract: Probabilities of causation (PoCs)…
  • Tian and Pearl first derived theoretically sharp bounds for binary PoC…
  • Mueller et al.

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

SteerBench-Work: A Benchmark for Agent Steering at Action Boundaries

arXiv:2608.12654v1 Announce Type: new Abstract: Long-running LLM agents act through tools, and a single step can send an email, merge a pull request, or wire a payment.…

  • The steering decision is the pre-commit choice at that boundary: proce…
  • We introduce SteerBench-Work, an incident-anchored, bidirectional benc…
  • Release v2026-05 contains 106 scenarios anchored in public incidents, …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

@skills: Attention is all you have

arXiv:2608.12610v1 Announce Type: new Abstract: There are 56,804 public agent skills today, and teams write many more privately.…

  • The dominant delivery model is installation: once installed, a skill's…
  • This leaves the long tail with no practical path to use and forces tea…
  • We observe that installation bundles three separable functions: conten…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence

arXiv:2608.12645v1 Announce Type: new Abstract: LLM judges have become central infrastructure for model evaluations, online grading, and reward modeling.…

  • Judges are typically validated by accuracy on golden data, but accurac…
  • We introduce the \emph{Wiggle Framework}, a unified stress test for ep…
  • The framework decomposes judge robustness along three dimensions: Mech…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues

arXiv:2608.…

  • 12599v1 Announce Type: new Abstract: Multi-turn dialogues let users re…
  • No existing instrument measures this influence per clause, predicts it…
  • \sysname{} closes the three gaps through the model API alone: a contra…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

DiG-bench: Discovery in Games

arXiv:2608.12593v1 Announce Type: new Abstract: Discovery---formulating novel generalizations---is a central part of the scientific process.…

  • Despite its importance, there is a gap in the current AI benchmark lan…
  • To address this gap, we release a new benchmark: DiG-bench (Discovery …
  • DiG-bench consists of a set of 70 independent games.

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting

arXiv:2608.…

  • 12590v1 Announce Type: new Abstract: Thyroid ultrasound diagnosis requ…
  • We present ThyroidXAgent, a clinician-interactive agentic AI system th…
  • The system was developed using OpenThyroidDB, a multicentre, multitask…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Reasoning Jury: Multi-Model Consensus for Evaluating Reasoning Traces

arXiv:2608.…

  • 12585v1 Announce Type: new Abstract: Improving reasoning LLMs requires…
  • Additionally, surfacing reasoning mistakes that the model makes would …
  • Due to the difficulty of this complex task on long reasoning traces, s…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Trie Automata for Constrained Decoding over Large Finite Sets

arXiv:2608.…

  • 12574v1 Announce Type: new Abstract: Large language models increasingl…
  • Current constrained decoding systems handle this through general-purpo…
  • We introduce the trie automaton, a specialized mechanism that exploits…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

CAS: A Causal Attribution Score for Local and Global Explainable Artificial Intelligence

arXiv:2608.…

  • 12555v1 Announce Type: new Abstract: Predictive explanation methods at…
  • We introduce the Causal Attribution Score (CAS), a compact score archi…
  • CAS starts from an identified interventional coalition game, allocates…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

$\varepsilon$-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution

arXiv:2608.…

  • 12522v1 Announce Type: new Abstract: LLM-based program evolution syste…
  • We introduce $\varepsilon$-MemEvo, a framework for cross-task knowledg…
  • $\varepsilon$-MemEvo stores prior experience as task-agnostic tactic m…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Governed Persistent Memory: Source-Bound State Semantics and Fail-Closed Release for Long-Horizon Agents

arXiv:2608.…

  • 12476v1 Announce Type: new Abstract: Long-term agent memory is usually…
  • We introduce Governed Persistent Memory (GPM), an auditable bitemporal…
  • Five executable clauses cover ledger integrity, source binding, confli…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

MindMemOS: A Portable and Self-Evolving Memory Operating Layer for AI Agents

arXiv:2608.…

  • 12428v1 Announce Type: new Abstract: Memory is a core component of AI …
  • However, existing memory systems often remain fixed after development,…
  • We present MindMemOS, a portable and self-evolving memory operating la…

RSS 官方收录 · 可信分层展示

详情 原文 分享图