Skip to main content

XMT

短闻

信流 · 上滑连读 · 来源可核

今日 稍后 搜索 RSS
1 / 24
Aggregate arXiv cs.AI 人工智能 45″

SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation

arXiv:2609.…

  • 03806v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) ge…
  • Progress, however, is held back by the lack of domain-specific evaluat…
  • We introduce \textbf{\ours}, a human-aligned evaluation framework for …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Govern the Model, Not Only the Data: Storage, Circulation, and Learning in Creative AI

arXiv:2609.…

  • 03800v1 Announce Type: new Abstract: Federated learning is increasingl…
  • It borrows the vocabulary of the federated social web, yet inverts its…
  • We argue that federation is not in itself a remedy for extractive AI, …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Transfiver: Human-AI Co-Inference through a Shared Editable State

arXiv:2609.…

  • 03797v1 Announce Type: new Abstract: Long-term human-AI interaction is…
  • We introduce the TRANSparent Framework for Interactive, Verifiable, Ed…
  • Its central idea is that interaction-specific information is maintaine…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions

arXiv:2609.…

  • 03787v1 Announce Type: new Abstract: AI agents increasingly gather evi…
  • A final output alone cannot show which evidence, tool state, rule, aut…
  • We present DNative-Twin, a graph-native digital twin that records a co…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Rethinking World Models for Safety-Critical Embodied Systems

arXiv:2609.…

  • 03774v1 Announce Type: new Abstract: World models have progressed from…
  • However, high predictive likelihood and visual fidelity do not necessa…
  • This perspective identifies three structural mismatches in current wor…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate InfoQ 中文 人工智能 6″

千问办公上线首月用户数突破3000万,企业用户占比过半

点击查看原文>

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation

arXiv:2609.…

  • 03753v1 Announce Type: new Abstract: As large language models (LLMs) b…
  • We introduce SimSkill, a self-evolving agent built around the Simulati…
  • SimSkill identifies capability gaps, generates and solves environment-…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation

arXiv:2609.…

  • 03727v1 Announce Type: new Abstract: Large language model agents can p…
  • Proactive service moves the decision upstream: an agent must infer ser…
  • This survey gives an operational definition centered on initiative and…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Artificial Intelligence for Energy Optimization in Data Centers

arXiv:2609.03716v1 Announce Type: new Abstract: Data centers are increasingly optimized by artificial intelligence and, at the same time, increasingly loaded by it.…

  • The literature treats these as two unrelated problems: control studies…
  • We screen roughly 194 papers retrieved through a documented protocol, …
  • Of 28 primary control-oriented studies, 18 are validated in simulation…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Counterfactual Routing Using Integer Programming with Constraint Generation

arXiv:2609.03707v1 Announce Type: new Abstract: We present our submission to the IJCAI 2025 'Counterfactual Routing Competition' (CRC 25).…

  • The goal of the competition is to find counterfactual explanations for…
  • This requires deciding what the minimal changes to a road network woul…
  • This enables explanations such as "Your suggested route would indeed h…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study

arXiv:2609.03702v1 Announce Type: new Abstract: General-purpose code embeddings power tools for code search, classification, and retrieval.…

  • Compact transformer encoders for code typically rely on either human-w…
  • We empirically study an alternative: contrastive pretraining of small …
  • We benchmark this approach against pretraining-based baselines, genera…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate 雷锋网 人工智能 45″

具身智能落地的最后20%,藏在「云」里

今年的WAIC和WRC,具身智能展台的风向变了:不再执着于炫技,转向务实。去年的画风还是十八般武艺,同台竞技——机器人跳街舞、格斗秀、走猫步等;而今年则是走进真实场景:机器人开始被放进物流、工业、家庭等具体任务里。…

  • 过去一个多月,雷峰网一线走访了灵初智能、艾欧智能等具身智能公司,也看了各厂商在展台上的演示,一个感受越来越强烈:真实场景里的任务,远比Dem…
  • 在走访和交流中,雷峰网也与多位云专家进行了交流。
  • 当被问及机器人落地、商业化时,一位云专家表示,具身智能存在明显的“长尾效应”:前60%、70%、80%的进展可能很快,但最后20%却需要花非…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Analysis of Prompt Engineering for Drug Toxicity Prediction

arXiv:2609.03635v1 Announce Type: new Abstract: Clinical trials in the UK can cost up to {\pounds}1.…

  • 3 million, with approximately 90% drug failure rate.
  • Toxicity is a major contributing factor in drug failure.
  • Testing is time and cost intensive.

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

A computable representation of the physical laboratory enables verifiable workflows

arXiv:2609.…

  • 03621v1 Announce Type: new Abstract: Making science computable require…
  • A computable representation of the physical laboratory is established …
  • It provides the physical-world counterpart to machine-readable knowled…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents

arXiv:2609.…

  • 03588v1 Announce Type: new Abstract: As LLMs increasingly act through …
  • We introduce KC-Bench, a controlled multi-turn benchmark for measuring…
  • Its 238 tasks are manually screened from more than 1,000 generated can…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

HalluPeer: A Taxonomy-driven Benchmark for Detecting Hallucinations in Scientific Peer Reviews

arXiv:2609.…

  • 03580v1 Announce Type: new Abstract: The growing scale of academic pee…
  • Existing hallucination benchmarks are not designed for peer review, wh…
  • We introduce HalluPeer, a benchmark for detecting hallucinations in sc…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

The Attention Triangle in Audio-Video Models

arXiv:2609.…

  • 03586v1 Announce Type: new Abstract: Audio-video diffusion models rely…
  • We study these models by probing and analyzing the ``attention triangl…
  • Our analysis reveals that routing along the audio-video edge is bidire…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate InfoQ 中文 人工智能 6″

让测试更加绿色可持续

点击查看原文>

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate InfoQ 中文 人工智能 6″

GPT-6 Astra 正式登场:烧了 10万块 GPU、多项跑分逼近满分,OpenAI 开启“AGI时代”

点击查看原文>

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

GPS-Bench: A Governance Policy Benchmark for Automating Policy Analysis

arXiv:2609.…

  • 03553v1 Announce Type: new Abstract: Policy analysis requires more tha…
  • LLM-based policy simulations model these processes at scale, but their…
  • We introduce GPS-Bench, an evidence-grounded benchmark for governance …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Dalek: A Constructive Agent Machine

arXiv:2609.…

  • 03546v1 Announce Type: new Abstract: We present Dalek, a closed machin…
  • The machine is built from three primitives---actors, messages, and cha…
  • Four obligations---a host boundary, a construction language, admissibl…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Feature Reconfiguration With Visual Prior for Medical Lesion Segmentation

arXiv:2609.03535v1 Announce Type: new Abstract: Lesion segmentation in medical images plays a critical role in clinical diagnosis and treatment planning.…

  • Despite significant advances, lesion segmentation remains challenging …
  • Existing encoder-decoder based methods mainly focus on enhancing featu…
  • However, they lack early prior guidance and feature reconfiguration du…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

NeoRed: A Knowledge-Logic-Alignment Multimodal Large Language Model for Neonatal Respiratory Disease Diagnosis

arXiv:2609.…

  • 03527v1 Announce Type: new Abstract: Neonatal respiratory diseases are…
  • Despite recent advances, existing Multimodal Large Language Models (ML…
  • To address these challenges, we collect two real-world clinical datase…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

CulturalMenuBench: Probing the Knowledge-Application Gap in Multimodal Culinary Reasoning

arXiv:2609.…

  • 03526v1 Announce Type: new Abstract: Multimodal language models achiev…
  • To probe this distinction, we introduce CulturalMenuBench, a benchmark…
  • Evaluating 12 models exposes a substantial knowledge-application gap: …

RSS 官方收录 · 可信分层展示

详情 原文 分享图