Skip to main content

XMT

短闻

信流 · 上滑连读 · 来源可核

今日 稍后 搜索 RSS

当前信源:arXiv cs.AI · 清除信源筛选

1 / 18
Aggregate arXiv cs.AI 人工智能 45″

InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk

arXiv:2608.11234v1 Announce Type: new Abstract: Managing modern computing infrastructure has become a steadily harder problem due to the ever-increasing complexity.…

  • Recent advances in AI agents create a timely opportunity to automate i…
  • We present InfraBench, a benchmark suite for evaluating AI agents on r…
  • Experiments with 15 agent-model configurations show that even the stro…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

LinearKV: One Cached State Suffices for Position-Independent Caching in Hybrid LLMs

arXiv:2608.11231v1 Announce Type: new Abstract: LLM serving is increasingly accelerated by position-independent caching (PIC).…

  • Existing PIC methods, however, are built for full-attention models, wh…
  • Hybrid LLMs break these primitives---they replace most attention layer…
  • This raises a natural question: can PIC benefit hybrid models, and wha…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

The Edge-based Contiguous p-median Problem with Connections to Logistics Districting

arXiv:2608.…

  • 11230v1 Announce Type: new Abstract: This paper introduces the edge-ba…
  • Two binary programming models are introduced, both of which incorporat…
  • The first model requires an exponential number of cut set-based constr…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended Version)

arXiv:2608.…

  • 11229v1 Announce Type: new Abstract: Comparative feedback, asking peop…
  • Preference-based reward learning typically casts the human teacher as …
  • We argue this forfeits the teacher's defining advantage: knowledge of …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Forecasting Side Effects of Activation Steering

arXiv:2608.…

  • 11227v1 Announce Type: new Abstract: Activation steering modifies a la…
  • While effective, steering often produces unintended side effects on ot…
  • We therefore ask: can these side effects be forecasted before steering…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Cutting AI Datacenter Energy with Reinforcement Learning: Measured Power Control of LLM Training from One GPU to the Fleet

arXiv:2608.…

  • 11226v1 Announce Type: new Abstract: Reinforcement-learning post-train…
  • We instrument GRPO training with half-second power telemetry at 7B, 14…
  • Against the full 500-step 7B trace, the controller cuts power-limit vi…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Identity from the Outside: A Conceptual Framework and Research Program for AI Personality Clones

arXiv:2608.11225v1 Announce Type: new Abstract: AI "personality clones" force a re-examination of personal identity in operational terms.…

  • Setting aside the hard problem of consciousness, we approach identity …
  • We distinguish three criteria that "identity" conflates: fidelity to a…
  • We propose a six-term factorization of observed identity (substrate, d…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Harnessing agent memory to build lifelong AI partners for materials scientists

arXiv:2608.…

  • 11224v1 Announce Type: new Abstract: Materials research advances throu…
  • This experience is essential for reproducibility and knowledge transfe…
  • Here we argue that a lifelong AI partner for materials science can be …

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

A Conceptual Framework for Refining Influence Knowledge from Simulation Evidence in Cyber-Physical Systems

arXiv:2608.…

  • 11221v1 Announce Type: new Abstract: Cyber-physical systems (CPS) are …
  • The behaviour of these systems emerges from the interaction between th…
  • Simulation and co-simulation have become essential approaches for anal…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

LLMs in Process Diagram Engineering: From Optimal PFDs to Validated P&IDs

arXiv:2608.…

  • 11220v1 Announce Type: new Abstract: Nowadays, the creation of a proce…
  • Applying artificial intelligence in the task could potentially lead no…
  • This research presents P&ID Pilot - a practical end-to-end AI pipeline…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

From Monolithic to Modular: Segment-level Automatic Prompt Optimization

arXiv:2608.11219v1 Announce Type: new Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others.…

  • We present SAPO, a segment-level APO method that decomposes prompts in…
  • The optimization loop uses one LLM with static meta-prompts and struct…
  • We describe a train/validation protocol and a two-stage generation pro…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

MaSRead: Content-Addressed Reading of Replicated Latent Stores

arXiv:2608.11218v1 Announce Type: new Abstract: Independent agents that reason in latent space can share computed state as key-value cache fragments rather than text.…

  • Merged by a conflict-free replicated data type, these fragments form a…
  • Yet a later query, unknown at encode time, cannot reliably read the me…
  • MaSRead addresses the read to content.

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research

arXiv:2608.…

  • 11216v1 Announce Type: new Abstract: World modeling is an unsettled fi…
  • This makes it an ideal testbed for AI coding agents acting as autonomo…
  • We introduce AutoWorldModel-Bench, a closed-loop benchmark in which fr…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop

arXiv:2608.…

  • 11215v1 Announce Type: new Abstract: Simulating societies of many larg…
  • We turn a statistical-physics observation into a method: replace each …
  • Whether this works is decided before the simulation runs, chiefly by w…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Damage in Quantized Mixture-of-Experts

arXiv:2608.…

  • 11212v1 Announce Type: new Abstract: Top-k Mixture-of-Experts (MoE) ro…
  • This paper proposes no new mitigation; it supplies a causal apparatus,…
  • A four-run apparatus prices the route-mediated fraction (RMF) of quant…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

A Forced-Structure Reduction and Verifiable Bounds for Conway's 99-Graph

arXiv:2608.11211v1 Announce Type: new Abstract: Conway's 99-graph problem asks whether a strongly regular graph with parameters $\mathrm{srg}(99,14,1,2)$ exists.…

  • We report a systematic, fully reproducible attack by an autonomous AI …
  • Our verifiable contributions are: (1) an exhaustive proof that no circ…
  • 0\%$ of the constraints ($33$ of $49$ difference-classes), with the sa…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration

arXiv:2608.11210v1 Announce Type: new Abstract: Bayesian calibration of process-based models requires a prior distribution for each model parameter.…

  • Despite decades of methodological work, researchers almost always fall…
  • The main reason is that building informative priors from scientific li…
  • We present \textbf{Distribird}, an agentic web application that automa…

RSS 官方收录 · 可信分层展示

详情 原文 分享图
Aggregate arXiv cs.AI 人工智能 45″

Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes

arXiv:2608.…

  • 11207v1 Announce Type: new Abstract: When two LLM agents with structur…
  • This paper asks whether a control-theoretic governance layer can subst…
  • The Experience Orchestrator (EO) addresses this in a simulated financi…

RSS 官方收录 · 可信分层展示

详情 原文 分享图