微信内可能无法直接打开本站。请点右上角 ··· → 在浏览器打开,或复制链接后用系统浏览器访问。
XMT
信流 · 上滑连读 · 来源可核
STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation
arXiv:2609.…
- 03874v1 Announce Type: new Abstract: Retrieval Augmented Generation (R…
- LLMs are improving at handling long context, but still suffer from "lo…
- Thus, precise and accurate retrieval is important.
RSS 官方收录 · 可信分层展示
Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations
arXiv:2609.03860v1 Announce Type: new Abstract: Retail supply chain operations rely on coupled decision modules that must adapt as requirements evolve.…
- LLMs offer a natural-language interface for this task, but existing me…
- Extending them to heterogeneous decision pipelines is challenging beca…
- We formulate requirement-driven adaptation as the joint selection of a…
RSS 官方收录 · 可信分层展示
CauseCollab: Causal Unified and Modality-Agnostic Network for Heterogeneous Collaborative Perception
arXiv:2609.…
- 03818v1 Announce Type: new Abstract: Collaborative perception enhances…
- Recent protocol-based two-stage methods alleviate this problem by mapp…
- To address this issue, we propose CauseCollab, a causal unified and mo…
RSS 官方收录 · 可信分层展示
SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation
arXiv:2609.…
- 03806v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) ge…
- Progress, however, is held back by the lack of domain-specific evaluat…
- We introduce \textbf{\ours}, a human-aligned evaluation framework for …
RSS 官方收录 · 可信分层展示
Govern the Model, Not Only the Data: Storage, Circulation, and Learning in Creative AI
arXiv:2609.…
- 03800v1 Announce Type: new Abstract: Federated learning is increasingl…
- It borrows the vocabulary of the federated social web, yet inverts its…
- We argue that federation is not in itself a remedy for extractive AI, …
RSS 官方收录 · 可信分层展示
Transfiver: Human-AI Co-Inference through a Shared Editable State
arXiv:2609.…
- 03797v1 Announce Type: new Abstract: Long-term human-AI interaction is…
- We introduce the TRANSparent Framework for Interactive, Verifiable, Ed…
- Its central idea is that interaction-specific information is maintaine…
RSS 官方收录 · 可信分层展示
DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions
arXiv:2609.…
- 03787v1 Announce Type: new Abstract: AI agents increasingly gather evi…
- A final output alone cannot show which evidence, tool state, rule, aut…
- We present DNative-Twin, a graph-native digital twin that records a co…
RSS 官方收录 · 可信分层展示
Rethinking World Models for Safety-Critical Embodied Systems
arXiv:2609.…
- 03774v1 Announce Type: new Abstract: World models have progressed from…
- However, high predictive likelihood and visual fidelity do not necessa…
- This perspective identifies three structural mismatches in current wor…
RSS 官方收录 · 可信分层展示
SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation
arXiv:2609.…
- 03753v1 Announce Type: new Abstract: As large language models (LLMs) b…
- We introduce SimSkill, a self-evolving agent built around the Simulati…
- SimSkill identifies capability gaps, generates and solves environment-…
RSS 官方收录 · 可信分层展示
Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation
arXiv:2609.…
- 03727v1 Announce Type: new Abstract: Large language model agents can p…
- Proactive service moves the decision upstream: an agent must infer ser…
- This survey gives an operational definition centered on initiative and…
RSS 官方收录 · 可信分层展示
Artificial Intelligence for Energy Optimization in Data Centers
arXiv:2609.03716v1 Announce Type: new Abstract: Data centers are increasingly optimized by artificial intelligence and, at the same time, increasingly loaded by it.…
- The literature treats these as two unrelated problems: control studies…
- We screen roughly 194 papers retrieved through a documented protocol, …
- Of 28 primary control-oriented studies, 18 are validated in simulation…
RSS 官方收录 · 可信分层展示
Counterfactual Routing Using Integer Programming with Constraint Generation
arXiv:2609.03707v1 Announce Type: new Abstract: We present our submission to the IJCAI 2025 'Counterfactual Routing Competition' (CRC 25).…
- The goal of the competition is to find counterfactual explanations for…
- This requires deciding what the minimal changes to a road network woul…
- This enables explanations such as "Your suggested route would indeed h…
RSS 官方收录 · 可信分层展示
Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study
arXiv:2609.03702v1 Announce Type: new Abstract: General-purpose code embeddings power tools for code search, classification, and retrieval.…
- Compact transformer encoders for code typically rely on either human-w…
- We empirically study an alternative: contrastive pretraining of small …
- We benchmark this approach against pretraining-based baselines, genera…
RSS 官方收录 · 可信分层展示
具身智能落地的最后20%,藏在「云」里
今年的WAIC和WRC,具身智能展台的风向变了:不再执着于炫技,转向务实。去年的画风还是十八般武艺,同台竞技——机器人跳街舞、格斗秀、走猫步等;而今年则是走进真实场景:机器人开始被放进物流、工业、家庭等具体任务里。…
- 过去一个多月,雷峰网一线走访了灵初智能、艾欧智能等具身智能公司,也看了各厂商在展台上的演示,一个感受越来越强烈:真实场景里的任务,远比Dem…
- 在走访和交流中,雷峰网也与多位云专家进行了交流。
- 当被问及机器人落地、商业化时,一位云专家表示,具身智能存在明显的“长尾效应”:前60%、70%、80%的进展可能很快,但最后20%却需要花非…
RSS 官方收录 · 可信分层展示
Analysis of Prompt Engineering for Drug Toxicity Prediction
arXiv:2609.03635v1 Announce Type: new Abstract: Clinical trials in the UK can cost up to {\pounds}1.…
- 3 million, with approximately 90% drug failure rate.
- Toxicity is a major contributing factor in drug failure.
- Testing is time and cost intensive.
RSS 官方收录 · 可信分层展示
A computable representation of the physical laboratory enables verifiable workflows
arXiv:2609.…
- 03621v1 Announce Type: new Abstract: Making science computable require…
- A computable representation of the physical laboratory is established …
- It provides the physical-world counterpart to machine-readable knowled…
RSS 官方收录 · 可信分层展示
KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents
arXiv:2609.…
- 03588v1 Announce Type: new Abstract: As LLMs increasingly act through …
- We introduce KC-Bench, a controlled multi-turn benchmark for measuring…
- Its 238 tasks are manually screened from more than 1,000 generated can…
RSS 官方收录 · 可信分层展示
HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews
arXiv:2609.…
- 03580v1 Announce Type: new Abstract: The growing scale of academic pee…
- Existing hallucination benchmarks are not designed for peer review, wh…
- We introduce HalluPeer, a benchmark for detecting hallucinations in sc…
RSS 官方收录 · 可信分层展示
The Attention Triangle in Audio-Video Models
arXiv:2609.…
- 03586v1 Announce Type: new Abstract: Audio-video diffusion models rely…
- We study these models by probing and analyzing the ``attention triangl…
- Our analysis reveals that routing along the audio-video edge is bidire…
RSS 官方收录 · 可信分层展示